Search: FSDP2
5 result(s) on page 1
pytorch-fsdp2
Indexed by skills.sh from orchestra-research/ai-research-skills
quarantinedskills.sh- Registry
- skills.sh
- Category
- Mlops· inferred
- Version
- 1.0.0
- Author
- orchestra-research
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__pytorch-fsdp2pytorch-fsdp
Fully sharded data-parallel training for large models.
quarantinedlinuxmacosDistributed Trainingoptional- Registry
- optional
- Category
- MLOps
- Version
- 1.0.0
- Author
- Orchestra Research
- License
- MIT
curl -s /v1/skills/community__axehub:optional__optional__pytorch-fsdptorchtitan
Pretrain LLMs at scale with PyTorch 4D parallelism.
quarantinedlinuxmacosModel Architectureoptional- Registry
- optional
- Category
- MLOps
- Version
- 1.0.0
- Author
- Orchestra Research
- License
- MIT
curl -s /v1/skills/community__axehub:optional__optional__torchtitannemo-automodel-distributed-training
Guide for selecting and configuring distributed training strategies in NeMo AutoModel, including FSDP2, Megatron FSDP, DDP, and parallelism settings.
quarantinedClawHub- Registry
- ClawHub
- Category
- Training Ai· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__nemo-automodel-distributed-trainingnemo-automodel-distributed-training
Guide for selecting and configuring distributed training strategies in NeMo AutoModel, including FSDP2, Megatron FSDP, DDP, and parallelism settings.
quarantinedNVIDIA- Registry
- NVIDIA
- Category
- Training AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__nemo-automodel-distributed-training