Projects in Awesome Lists tagged with fsdp
A curated list of projects in awesome lists tagged with fsdp .
https://github.com/lambdalabsml/distributed-training-guide
Best practices & guides on how to write distributed pytorch training code
cluster cuda deepspeed distributed-training fsdp gpu gpu-cluster kuberentes lambdalabs mpi nccl pytorch sharding slurm
Last synced: 16 May 2025
https://github.com/meta-pytorch/torchft
Fault tolerance for PyTorch (HSDP, LocalSGD, DiLoCo, Streaming DiLoCo)
consensus diloco distributed-systems fault-tolerance fsdp hsdp localsgd ml pytorch torch training
Last synced: 05 Oct 2025
https://github.com/LambdaLabsML/distributed-training-guide
Best practices & guides on how to write distributed pytorch training code
cluster cuda deepspeed distributed-training fsdp gpu gpu-cluster kuberentes lambdalabs mpi nccl pytorch sharding slurm
Last synced: 08 Mar 2025
https://github.com/tsiendragon/qwen-image-finetune
Repo for Qwen Image Finetune
diffusion-models flux-kontext fsdp image-edit-model image-to-image lora peft-fine-tuning-llm qwen-image-edit qwen-image-edit-2509
Last synced: 04 Apr 2026
https://github.com/saforem2/ezpz
Write once, run anywhere; ezpz 🍋
ai-tools deepspeed distributed-training fsdp launcher machine-learning mpi mpi4py parallelism python pytorch rich slurm torch training
Last synced: 26 Jun 2026
https://github.com/gurpreetkaurjethra/meta-llama3-genai-usecases-end-to-end-implementation-guides
META LLAMA3 GENAI Real World UseCases End To End Implementation Guide
chromadb fine-tuning fsdp generativeai huggingface langchain-python llama3 llama3-70b-8192 llama3-finetune llama3-meta-ai llama3-prompts llama3-rag ollama prompt-tuning pytorch qlora rag sagemaker streamlit
Last synced: 06 Oct 2025
https://github.com/abhilash1910/framework-optimization
Framework, Model & Kernel Optimizations for Distributed Deep Learning - Data Hack Summit
codegen ddp deepspeed fsdp inductor pipelineparallel pytorch tensorparallel triton
Last synced: 19 May 2026
https://github.com/hyunnnchoi/google-t5-fsdp-kubeflow
A foundational repository for setting up distributed training jobs using Kubeflow and PyTorch FSDP.
distributed-deep-learning fsdp kubeflow pytorch
Last synced: 26 Apr 2026
https://github.com/tsugiai/tsugi
Unified developer surface for TsugiCinema's open-source distributed-training SDKs. pip install tsugi.
distributed-training fsdp llm-training lora machine-learning open-source peft pytorch
Last synced: 29 May 2026
https://github.com/hrolive/large-language-models-on-supercomputers
Comprehensive exploration of LLMs, including cutting-edge techniques and tools such as parameter-efficient fine-tuning (PEFT), quantization, zero redundancy optimizers (ZeRO), fully sharded data parallelism (FSDP), DeepSpeed, and Huggingface accelerate.
deepspeed evaluation-metrics fsdp high-performance-computing hpc huggingface huggingface-transformers jupyter llm llm-inference llm-training monitoring peft python quantization slurm tokenization transformer unsloth
Last synced: 18 May 2026
https://github.com/tsugiai/tsugi-mend
Cross-rack distributed-training reducer for PyTorch. Apache-2.0, patent-independent. Part of the unified pip install tsugi surface.
cross-rack diloco distributed-training fault-tolerance fsdp gpu llm-training pytorch
Last synced: 29 May 2026
https://github.com/debnsuma/ray-for-developers
A comprehensive hands-on guide to building production-grade distributed applications with Ray - from distributed training and multimodal data processing to inference and reinforcement learning.
ddp deep-learning distributed-computing distributed-training fsdp machine-learning mlops model-serving multimodal pytorch ray
Last synced: 18 Jun 2026
https://github.com/shreyansh26/wordle-solver
Training Qwen3 to solve Wordle using SFT and GRPO
fsdp grpo llm qwen3 rft rl sft tensor-parallelism wordle wordle-solver
Last synced: 18 Apr 2026