NVIDIA NeMo β€” End-to-End Flow

Data β†’ training β†’ parallelism β†’ post-training β†’ evaluation β†’ export β†’ serving, across the modular NeMo Framework repos.
βœ“ Verified against NVIDIA-NeMo GitHub repos + NVIDIA NIM docs Β· Sep 29, 2026

Sources (github.com/NVIDIA-NeMo, Apache-2.0): Curator Β· Automodel Β· Megatron-Bridge Β· RL Β· Evaluator Β· Export-Deploy Β· Speech (formerly NeMo) Β· org overview table. NVIDIA NIM is a separate NVIDIA product (developer.nvidia.com/nim), not a NeMo Framework repo β€” shown dashed.
Kubernetes layer: vendor-neutral. The object types shown (RayCluster, PyTorchJob or JobSet, Job, Deployment, Helm, NIM Operator) are common ways to run each stage, not NeMo requirements, and NeMo also runs on Slurm.
Simplifications: stage order is illustrative. Real workflows skip or repeat stages, and evaluation often happens throughout. Parallelism is part of training; it's shown separately here for clarity.