mirror of
https://github.com/NVIDIA/Model-Optimizer.git
synced 2026-10-02 03:14:52 +08:00
Train -> serve -> benchmark in one YAML, so a handover has a single reproducible chain rather than artifacts stitched from separate examples. Trains from the full Spec-Decoding-Dataset-v2 with a bounded max_steps, so nothing needs staging first. Signed-off-by: Ye Yu <yeyu@nvidia.com> Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Signed-off-by: Ye Yu <yeyu@nvidia.com>