Files
Chad Voegele 28dc117594 Add Day 0 Sub-agent Roles (#2006)
### What does this PR do?

Type of change: new feature

Adding sub-agent definition for Day 0 workflows.

I'm adding subagents as an alternative, while we evaluate which is
better. They are duplicated because Claude & Codex have different
formats & different locations. Deriving from a shared vendor-neutral
format at build-time is too complicated.

### Usage

Tell your agent "Use the modelopt_model_quantizer agent to quantize the
model"

### Testing

Ran a trial using Qwen-2.5, Muse Glimmer, Qwen-2.8, and GLM-5.3-Flash.

### Before your PR is "*Ready for review*"

Make sure you read and follow [Contributor
guidelines](https://github.com/NVIDIA/Model-Optimizer/blob/main/CONTRIBUTING.md)
and your commits are signed (`git commit -s -S`).

Make sure you read and follow the [Security Best
Practices](https://github.com/NVIDIA/Model-Optimizer/blob/main/SECURITY.md#security-coding-practices-for-contributors)
(e.g. avoiding hardcoded `trust_remote_code=True`, `torch.load(...,
weights_only=False)`, `pickle`, etc.).

- Is this change backward compatible?: ✅
- If you copied code from any other sources or added a new PIP
dependency, did you follow guidance in `CONTRIBUTING.md`: ✅
- Did you write any new necessary tests?: ✅
- Did you update
[Changelog](https://github.com/NVIDIA/Model-Optimizer/blob/main/CHANGELOG.rst)?:
✅ / ❌ / N/A <!--- Only for new features, API changes, critical bug fixes
or backward incompatible changes. -->
- Did you get Claude approval on this PR?: ✅ / ❌ / N/A <!--- Run
`/claude review`. NVIDIA org members can self-trigger for complex
changes; orthogonal to CodeRabbit. -->

### Additional Information
See design doc.


<!-- This is an auto-generated comment: release notes by coderabbit.ai
-->
## Summary by CodeRabbit

## New Features
- Added specialized Model Optimizer agents for downloading,
quantization, recipe search, evaluation, deployment, and performance
benchmarking.
- Added coordinated Codex and Claude Code access with validation
requirements, artifact preservation, and concise handoffs.
- Improved skill discovery across plugin and repository installations.

## Documentation
- Clarified agent discovery, layout, and canonical editing locations.
- Updated configuration guidance for supported agent definitions.

## Tests
- Added synchronization checks for agent definitions and links.
- Improved test compatibility for Python versions below 3.11.
<!-- end of auto-generated comment: release notes by coderabbit.ai -->

---------

Signed-off-by: Chad Voegele <cvoegele@nvidia.com>
2026-09-10 17:30:37 +00:00

82 lines
1.3 KiB
Plaintext

# Byte-compiled / optimized / DLL files
**/__pycache__
**.py[cod]
**$py.class
# C, CPP extensions
*.so
*.so.lock
**.rendered.*.cpp
**.rendered.*.o
# Distribution / packaging
build/
dist/
*.egg-info/
# Unit test / coverage reports
htmlcov/
.coverage
.coverage.*
coverage.xml
.pytest_cache/
# Sphinx documentation
docs/build
docs/source/reference/generated
# Jupyter Notebook
**/.ipynb_checkpoints
# Agent session workspaces (see plugins/modelopt/skills/common/workspace-management.md)
workspaces/
# Environments
.env*
.venv
env/
venv/
# Linters
**/.mypy_cache
**/.ruff_cache
# Vscode
.vscode/*
!.vscode/settings.json
!.vscode/extensions.json
# Mac stuff
**/.DS_Store
# Ignore experiment checkpoints
**.pt
**.pth.tar
**.pth
**.pb
**.onnx
**.ckpt
**.safetensors
**.bin
**.pkl
**.pickle
**.tar.gz
# Ignore claude local settings and runtime agent state
.claude/settings.local.json
.claude/agents/*
!.claude/agents/modelopt-model-deployer.md
!.claude/agents/modelopt-model-downloader.md
!.claude/agents/modelopt-model-evaluator.md
!.claude/agents/modelopt-model-performance-benchmarker.md
!.claude/agents/modelopt-model-quantize-recipe-searcher.md
!.claude/agents/modelopt-model-quantizer.md
CLAUDE.local.md
AGENTS.override.md
# Ignore SonarQube analysis
.sonar/
# Claude Code runtime lock (ephemeral process state — never commit)
.claude/scheduled_tasks.lock