NVIDIA Nemotron 3 Super is a 120B open MoE model activating just 12B parameters to deliver maximum compute efficiency and accuracy for complex multi-agent applications.
ollama run nemotron-3-super
策展 + Ollama Library 全量 + HF GGUF 热门 合并目录:参数/显存、许可、 Hugging Face / GGUF / Ollama / ModelScope 下载,并跳转 本地算力计算器 。HF 全库十万+——本频道做可跑清单、深链与算力账,不镜像权重。 同步于 2026-08-03T07:33:50Z。
要「全部开源」请用下列权威源检索;本站策展负责选型 + 部署路径 + 算力账。
匹配 9 · 展示 9 · 点「设为 A/B」后底部一键对比 · 本站不托管权重
NVIDIA Nemotron 3 Super is a 120B open MoE model activating just 12B parameters to deliver maximum compute efficiency and accuracy for complex multi-agent applications.
ollama run nemotron-3-super
高效 MoE 向;开放权重适合推理栈。
更大 Nemotron;机房/多卡。
Llama-3.1-Nemotron-70B-Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM generated responses to user queries.
ollama run nemotron
Nemotron-3-Nano is a new Standard for Efficient, Open, and Intelligent Agentic Models, now updated with a 4B parameter count model.
ollama run nemotron-3-nano
A commercial-friendly small language model by NVIDIA optimized for roleplay, RAG QA, and function calling.
ollama run nemotron-mini
NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows.
ollama run nemotron3
An open 30B MoE model from NVIDIA with 3B activated parameters that delivers strong reasoning and agentic capabilities.
ollama run nemotron-cascade-2
NVIDIA Nemotron 3 Ultra is built for high-throughput reasoning and long-running agent workflows.
ollama run nemotron-3-ultra
ollama run,不二次托管权重。Clavue CLI / imux IDE / clavue-2.1
CLI / GUI Agent 运行时、设备登录、会员额度与 OpenAI 兼容 API(api.clavue.com)。开发脚本、CI 与本地工具一条链路。
打开 Clavue ↗Clavue 平台的原生 IDE:多 Agent 分屏、Ghostty 级终端、Agent Chat、浏览器自动化与 Supervisor——不是又一个 Electron 壳。
了解 imux ↗旗舰模型 clavue-2.1(及 fast / pro / rev):适合复杂推理与长程 Agent 循环。Chat、imux、API 共用会员额度。
查看模型 ↗