← Browse
distributed-training
HermeticOrmus ★ 0
Provides expert guidance on scaling PyTorch training across multiple GPUs and nodes using DDP, FSDP, DeepSpeed ZeRO, Megatron-LM, and mixed precision strategies. Includes code patterns, launch configurations for torchrun and SLURM, debugging workflows, and memory optimization techniques like gradient checkpointing.
Install
> /plugin marketplace add HermeticOrmus/LibreMLOps-Claude-Code/tree/HEAD/plugins/distributed-training
> /plugin install distributed-training
Source: https://github.com/HermeticOrmus/LibreMLOps-Claude-Code/tree/HEAD/plugins/distributed-training
What it's made of
1 command · 1 agent · 1 skill
- Commands
- 1
- Agents
- 1
- Skills
- 1
- MCP servers
- 0
- Hooks
- 0
What it needs & plugs into
- API keys
- none
- Paid services
- none detected
- External tools
- none
- Talks to
- nothing external detected
Analyzed . Facts extracted from the plugin's files. Prose generated by claude-haiku-4-5-20251001.
Is this your plugin?
Claim it to keep the card accurate and enter the weekly contest. Requires signing in as the GitHub owner (HermeticOrmus).