mirror of
https://github.com/NVIDIA/Model-Optimizer.git
synced 2026-10-02 03:14:52 +08:00
- export_tensorrt_llm_checkpoint: when decoder_type cannot be detected, raise a clear ValueError unless the model has config.architectures for the generic decoder path. A Megatron-Core caller that omits decoder_type otherwise failed late with KeyError: 'unknown:GPTModel'. - requantize_resmooth_fused_llm_layers: detect Nemotron VL from config.architectures, like is_multimodal_model and is_nemotron_vl, since model_type is free-form for these remote-code models. - Drop the stale "(t5/bart/whisper)" from the layerwise-export refusal label; is_enc_dec also covers mt5, umt5 and mbart. - Shorten the CHANGELOG entry to two sentences. Signed-off-by: Shengliang Xu <shengliangx@nvidia.com>