NVIDIA/Model-Optimizer/accessing-mlflow A Agent Skill Active
Query and browse evaluation results stored in MLflow.
Browse SKILL.md-based agent skills for Claude Code, Codex, Cursor, Gemini CLI and more.
Query and browse evaluation results stored in MLflow.
Run commands inside a remote Docker container via the file-based command relay (tools/debugger).
Serve a quantized or unquantized LLM checkpoint as an OpenAI-compatible API endpoint using vLLM, SGLang, or TRT-LLM.
Evaluates accuracy of quantized or unquantized LLMs using NeMo Evaluator Launcher (NEL).
Run, monitor, analyze, and debug LLM evaluations via nemo-evaluator-launcher.
Monitor submitted jobs (PTQ, evaluation, deployment) on SLURM clusters.
This skill should be used when the user asks to "quantize a model", "run PTQ", "post-training quantization", "NVFP4 quantization", "FP8 quantization", "INT8 quantization", "INT4 AW...
Cherry-pick merged PRs labeled for a release branch into that branch, then open a PR and apply the cherry-pick-done label.