mirror of
https://github.com/NVIDIA/Model-Optimizer.git
synced 2026-10-02 03:14:52 +08:00
### What does this PR do? Type of change: new feature Adding sub-agent definition for Day 0 workflows. I'm adding subagents as an alternative, while we evaluate which is better. They are duplicated because Claude & Codex have different formats & different locations. Deriving from a shared vendor-neutral format at build-time is too complicated. ### Usage Tell your agent "Use the modelopt_model_quantizer agent to quantize the model" ### Testing Ran a trial using Qwen-2.5, Muse Glimmer, Qwen-2.8, and GLM-5.3-Flash. ### Before your PR is "*Ready for review*" Make sure you read and follow [Contributor guidelines](https://github.com/NVIDIA/Model-Optimizer/blob/main/CONTRIBUTING.md) and your commits are signed (`git commit -s -S`). Make sure you read and follow the [Security Best Practices](https://github.com/NVIDIA/Model-Optimizer/blob/main/SECURITY.md#security-coding-practices-for-contributors) (e.g. avoiding hardcoded `trust_remote_code=True`, `torch.load(..., weights_only=False)`, `pickle`, etc.). - Is this change backward compatible?: ✅ - If you copied code from any other sources or added a new PIP dependency, did you follow guidance in `CONTRIBUTING.md`: ✅ - Did you write any new necessary tests?: ✅ - Did you update [Changelog](https://github.com/NVIDIA/Model-Optimizer/blob/main/CHANGELOG.rst)?: ✅ / ❌ / N/A <!--- Only for new features, API changes, critical bug fixes or backward incompatible changes. --> - Did you get Claude approval on this PR?: ✅ / ❌ / N/A <!--- Run `/claude review`. NVIDIA org members can self-trigger for complex changes; orthogonal to CodeRabbit. --> ### Additional Information See design doc. <!-- This is an auto-generated comment: release notes by coderabbit.ai --> ## Summary by CodeRabbit ## New Features - Added specialized Model Optimizer agents for downloading, quantization, recipe search, evaluation, deployment, and performance benchmarking. - Added coordinated Codex and Claude Code access with validation requirements, artifact preservation, and concise handoffs. - Improved skill discovery across plugin and repository installations. ## Documentation - Clarified agent discovery, layout, and canonical editing locations. - Updated configuration guidance for supported agent definitions. ## Tests - Added synchronization checks for agent definitions and links. - Improved test compatibility for Python versions below 3.11. <!-- end of auto-generated comment: release notes by coderabbit.ai --> --------- Signed-off-by: Chad Voegele <cvoegele@nvidia.com>
82 lines
1.3 KiB
Plaintext
82 lines
1.3 KiB
Plaintext
# Byte-compiled / optimized / DLL files
|
|
**/__pycache__
|
|
**.py[cod]
|
|
**$py.class
|
|
|
|
# C, CPP extensions
|
|
*.so
|
|
*.so.lock
|
|
**.rendered.*.cpp
|
|
**.rendered.*.o
|
|
|
|
# Distribution / packaging
|
|
build/
|
|
dist/
|
|
*.egg-info/
|
|
|
|
# Unit test / coverage reports
|
|
htmlcov/
|
|
.coverage
|
|
.coverage.*
|
|
coverage.xml
|
|
.pytest_cache/
|
|
|
|
# Sphinx documentation
|
|
docs/build
|
|
docs/source/reference/generated
|
|
|
|
# Jupyter Notebook
|
|
**/.ipynb_checkpoints
|
|
|
|
# Agent session workspaces (see plugins/modelopt/skills/common/workspace-management.md)
|
|
workspaces/
|
|
|
|
# Environments
|
|
.env*
|
|
.venv
|
|
env/
|
|
venv/
|
|
|
|
# Linters
|
|
**/.mypy_cache
|
|
**/.ruff_cache
|
|
|
|
# Vscode
|
|
.vscode/*
|
|
!.vscode/settings.json
|
|
!.vscode/extensions.json
|
|
|
|
# Mac stuff
|
|
**/.DS_Store
|
|
|
|
# Ignore experiment checkpoints
|
|
**.pt
|
|
**.pth.tar
|
|
**.pth
|
|
**.pb
|
|
**.onnx
|
|
**.ckpt
|
|
**.safetensors
|
|
**.bin
|
|
**.pkl
|
|
**.pickle
|
|
**.tar.gz
|
|
|
|
# Ignore claude local settings and runtime agent state
|
|
.claude/settings.local.json
|
|
.claude/agents/*
|
|
!.claude/agents/modelopt-model-deployer.md
|
|
!.claude/agents/modelopt-model-downloader.md
|
|
!.claude/agents/modelopt-model-evaluator.md
|
|
!.claude/agents/modelopt-model-performance-benchmarker.md
|
|
!.claude/agents/modelopt-model-quantize-recipe-searcher.md
|
|
!.claude/agents/modelopt-model-quantizer.md
|
|
CLAUDE.local.md
|
|
AGENTS.override.md
|
|
|
|
# Ignore SonarQube analysis
|
|
.sonar/
|
|
|
|
# Claude Code runtime lock (ephemeral process state — never commit)
|
|
.claude/scheduled_tasks.lock
|