Home › Model-Optimizer

Model-Optimizer

8 MCP servers and agent skills in the Model-Optimizer category, ranked by quality score — 8 results.

Agent Skill Active

NVIDIA/Model-Optimizer/accessing-mlflow

Query and browse evaluation results stored in MLflow.

2.8k Python Updated today Score 84
Agent Skill Active

NVIDIA/Model-Optimizer/debug

Run commands inside a remote Docker container via the file-based command relay (tools/debugger).

2.8k Python Updated today Score 84
Agent Skill Active

NVIDIA/Model-Optimizer/deployment

Serve a quantized or unquantized LLM checkpoint as an OpenAI-compatible API endpoint using vLLM, SGLang, or TRT-LLM.

2.8k Python Updated today Score 84
Agent Skill Active

NVIDIA/Model-Optimizer/evaluation

Evaluates accuracy of quantized or unquantized LLMs using NeMo Evaluator Launcher (NEL).

2.8k Python Updated today Score 84
Agent Skill Active

NVIDIA/Model-Optimizer/launching-evals

Run, monitor, analyze, and debug LLM evaluations via nemo-evaluator-launcher.

2.8k Python Updated today Score 84
Agent Skill Active

NVIDIA/Model-Optimizer/monitor

Monitor submitted jobs (PTQ, evaluation, deployment) on SLURM clusters.

2.8k Python Updated today Score 84
Agent Skill Active

NVIDIA/Model-Optimizer/ptq

This skill should be used when the user asks to "quantize a model", "run PTQ", "post-training quantization", "NVFP4 quantization", "FP8 quantization", "INT8 quantization", "INT4 AW...

2.8k Python Updated today Score 84
Agent Skill Active

NVIDIA/Model-Optimizer/release-cherry-pick

Cherry-pick merged PRs labeled for a release branch into that branch, then open a PR and apply the cherry-pick-done label.

2.8k Python Updated today Score 84