Open Models · Local Deploy

全球开源模型频道

策展 + Ollama Library 全量 + HF GGUF 热门 合并目录:参数/显存、许可、 Hugging Face / GGUF / Ollama / ModelScope 下载,并跳转 本地算力计算器 。HF 全库十万+——本频道做可跑清单、深链与算力账,不镜像权重。 同步于 2026-08-03T07:33:50Z

合并总量325
策展精选42
同步条数303
本地优先318
含 Ollama238
含 HF/GGUF192

全库入口(全球)

要「全部开源」请用下列权威源检索;本站策展负责选型 + 部署路径 + 算力账

按家族

我的显卡一键筛:

匹配 48 · 展示 40 · 点「设为 A/B」后底部一键对比 · 本站不托管权重

qwen2.5-coder

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 20M pulls

Q4 ~4GB舒适 8GB

The latest series of Code-Specific Qwen models, with significant improvements in code generation, code reasoning, and code fixing.

代码推理Agent对话
ollama run qwen2.5-coder
qwen

Alibaba · 0.5b/110b · unknown · 本地优先 · ollama-library · 7.5M pulls

Q4 ~38.5GB舒适 49GB

Qwen 1.5 is a series of large language models by Alibaba Cloud spanning from 0.5B to 110B parameters

对话
ollama run qwen
qwen2

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 6.1M pulls

Q4 ~4GB舒适 8GB

Qwen2 is a new series of large language models from Alibaba group

Agent对话
ollama run qwen2
qwen3-vl

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 5.0M pulls

Q4 ~6GB舒适 12GB

The most powerful vision-language model in the Qwen model family to date.

视觉推理Agent对话
ollama run qwen3-vl
qwen3.5

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 17M pulls

Q4 ~6GB舒适 12GB

Qwen 3.5 is a family of open-source multimodal models that delivers exceptional utility and performance.

视觉推理Agent对话
ollama run qwen3.5
qwen3.6

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 5.1M pulls

Q4 ~6GB舒适 12GB

Qwen3.6 delivers substantial upgrades in agentic coding and thinking preservation than previous Qwen models.

视觉代码推理Agent对话
ollama run qwen3.6
codeqwen

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 1.1M pulls

Q4 ~4GB舒适 8GB

CodeQwen1.5 is a large language model pretrained on a large amount of code data.

代码对话
ollama run codeqwen
cogito

Community · ~7B · unknown · 本地优先 · ollama-library · 2.1M pulls

Q4 ~4GB舒适 8GB

Cogito v1 Preview is a family of hybrid reasoning models by Deep Cogito that outperform the best available open models of the same size, including counterparts from LLaMA, DeepSeek, and Qwen across most standard benchmarks.

推理Agent对话
ollama run cogito
deepscaler

Community · 1.5b · unknown · 本地优先 · ollama-library · 1.3M pulls

Q4 ~4GB舒适 8GB

A fine-tuned version of Deepseek-R1-Distilled-Qwen-1.5B that surpasses the performance of OpenAI’s o1-preview with just 1.5B parameters on popular math evaluations.

推理数学对话
ollama run deepscaler
qwen2-math

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 1.1M pulls

Q4 ~4GB舒适 8GB

Qwen2 Math is a series of specialized math language models built upon the Qwen2 LLMs, which significantly outperforms the mathematical capabilities of open-source models and even closed-source models (e.g., GPT4o).

数学对话
ollama run qwen2-math
qwen2.5vl

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 3.7M pulls

Q4 ~6GB舒适 12GB

Flagship vision-language model of Qwen and also a significant leap from the previous Qwen2-VL.

视觉对话
ollama run qwen2.5vl
qwen3-coder-next

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 1.9M pulls

Q4 ~4GB舒适 8GB

Qwen3-Coder-Next is a coding-focused language model from Alibaba's Qwen team, optimized for agentic coding workflows and local development.

代码Agent对话
ollama run qwen3-coder-next
qwen3-embedding

Alibaba · ~1B · unknown · 本地优先 · ollama-library · 3.5M pulls

Q4 ~0.8GB舒适 2GB

Building upon the foundational models of the Qwen3 series, Qwen3 Embedding provides a comprehensive range of text embeddings models in various sizes

向量
ollama run qwen3-embedding
qwq

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 2.3M pulls

Q4 ~4GB舒适 8GB

QwQ is the reasoning model of the Qwen series.

推理Agent对话
ollama run qwq
qwen3-next

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 577K pulls

Q4 ~4GB舒适 8GB

The first installment in the Qwen3-Next series with strong performance in terms of both parameter efficiency and inference speed.

推理Agent对话
ollama run qwen3-next
smallthinker

Community · 3b · unknown · 本地优先 · ollama-library · 254K pulls

Q4 ~4GB舒适 8GB

A new small reasoning model fine-tuned from the Qwen 2.5 3B Instruct model.

推理对话
ollama run smallthinker

与竞品对齐的能力

  • ApX / VRAM 计算器类:本站 /tools/local-deploy 覆盖显存适配、多卡 TP、产出与 vs API 费用。
  • Ollama Library / HF Hub:每条模型直链下载与 ollama run,不二次托管权重。
  • Model Picker(按 GPU 筛模型):列表「显存上限」筛选 + 计算器预填。
  • 价格/中转对照:有 OpenRouter id 的可回 /models · /official-api
开源模型部署频道 · HF / Ollama / GGUF · GrokCode 倍率榜