Open Models · Local Deploy

全球开源模型频道

策展 + Ollama Library 全量 + HF GGUF 热门 合并目录:参数/显存、许可、 Hugging Face / GGUF / Ollama / ModelScope 下载,并跳转 本地算力计算器 。HF 全库十万+——本频道做可跑清单、深链与算力账,不镜像权重。 同步于 2026-08-03T07:33:50Z

合并总量325
策展精选42
同步条数303
本地优先318
含 Ollama238
含 HF/GGUF192

全库入口(全球)

要「全部开源」请用下列权威源检索;本站策展负责选型 + 部署路径 + 算力账

按家族

我的显卡一键筛:

匹配 138 · 展示 40 · 点「设为 A/B」后底部一键对比 · 本站不托管权重

gpt-oss

Community · ~7B · unknown · 本地优先 · ollama-library · 11M pulls

Q4 ~4GB舒适 8GB

OpenAI’s open-weight models designed for powerful reasoning, agentic tasks, and versatile developer use cases.

推理Agent对话
ollama run gpt-oss
aya

Community · ~7B · unknown · 本地优先 · ollama-library · 1.1M pulls

Q4 ~4GB舒适 8GB

Aya 23, released by Cohere, is a new family of state-of-the-art, multilingual models that support 23 languages.

对话多语
ollama run aya
cogito

Community · ~7B · unknown · 本地优先 · ollama-library · 2.1M pulls

Q4 ~4GB舒适 8GB

Cogito v1 Preview is a family of hybrid reasoning models by Deep Cogito that outperform the best available open models of the same size, including counterparts from LLaMA, DeepSeek, and Qwen across most standard benchmarks.

推理Agent对话
ollama run cogito
command-r

Community · ~7B · unknown · 本地优先 · ollama-library · 1.5M pulls

Q4 ~4GB舒适 8GB

Command R is a Large Language Model optimized for conversational interaction and long context tasks.

Agent对话
ollama run command-r
deepscaler

Community · 1.5b · unknown · 本地优先 · ollama-library · 1.3M pulls

Q4 ~4GB舒适 8GB

A fine-tuned version of Deepseek-R1-Distilled-Qwen-1.5B that surpasses the performance of OpenAI’s o1-preview with just 1.5B parameters on popular math evaluations.

推理数学对话
ollama run deepscaler
falcon

Community · ~7B · unknown · 本地优先 · ollama-library · 1.1M pulls

Q4 ~4GB舒适 8GB

A large language model built by the Technology Innovation Institute (TII) for use in summarization, text generation, and chat bots.

对话
ollama run falcon
falcon3

Community · 10b · unknown · 本地优先 · ollama-library · 2.6M pulls

Q4 ~8GB舒适 16GB

A family of efficient AI models under 10B parameters performant in science, math, and coding through innovative training techniques.

代码数学对话
ollama run falcon3
gemini-3-flash-preview

Community · ~7B · unknown · 本地优先 · ollama-library · 2.3M pulls

Q4 ~6GB舒适 12GB

Gemini 3 Flash offers frontier intelligence built for speed at a fraction of the cost.

视觉推理Agent对话
ollama run gemini-3-flash-preview
granite-code

Community · ~7B · unknown · 本地优先 · ollama-library · 1.5M pulls

Q4 ~4GB舒适 8GB

A family of open foundation models by IBM for Code Intelligence

代码对话
ollama run granite-code
granite3-dense

Community · 2b/8b · unknown · 本地优先 · ollama-library · 1.0M pulls

Q4 ~4.4GB舒适 8.8GB

The IBM Granite 2B and 8B models are designed to support tool-based use cases and support for retrieval augmented generation (RAG), streamlining code generation, translation and bug fixing.

代码Agent对话
ollama run granite3-dense
granite3.1-dense

Community · 2b/8b · unknown · 本地优先 · ollama-library · 1.0M pulls

Q4 ~4.4GB舒适 8.8GB

The IBM Granite 2B and 8B models are text-only dense LLMs trained on over 12 trillion tokens of data, demonstrated significant improvements over their predecessors in performance and speed in IBM’s initial testing.

Agent对话
ollama run granite3.1-dense
granite3.1-moe

Community · 1b/3b · unknown · 本地优先 · ollama-library · 3.0M pulls

Q4 ~4GB舒适 8GB

The IBM Granite 1B and 3B models are long-context mixture of experts (MoE) Granite models from IBM designed for low latency usage.

MoEAgent对话
ollama run granite3.1-moe
granite3.3

Community · 2b/8b · unknown · 本地优先 · ollama-library · 1.1M pulls

Q4 ~4.4GB舒适 8.8GB

IBM Granite 2B and 8B models are 128K context length language models that have been fine-tuned for improved reasoning and instruction-following capabilities.

推理Agent对话
ollama run granite3.3
granite4

Community · ~7B · unknown · 本地优先 · ollama-library · 1.4M pulls

Q4 ~4GB舒适 8GB

Granite 4 features improved instruction following (IF) and tool-calling capabilities, making them more effective in enterprise applications.

Agent对话
ollama run granite4
hermes3

Community · ~7B · unknown · 本地优先 · ollama-library · 1.5M pulls

Q4 ~4GB舒适 8GB

Hermes 3 is the latest version of the flagship Hermes series of LLMs by Nous Research

Agent对话
ollama run hermes3
lfm2

Community · 24b · unknown · 本地优先 · ollama-library · 1.1M pulls

Q4 ~14GB舒适 24GB

LFM2 is a family of hybrid models designed for on-device deployment. LFM2-24B-A2B is the largest model in the family, scaling the architecture to 24 billion parameters while keeping inference efficient.

Agent对话
ollama run lfm2
lfm2.5-thinking

Community · ~7B · unknown · 本地优先 · ollama-library · 1.2M pulls

Q4 ~4GB舒适 8GB

LFM2.5 is a new family of hybrid models designed for on-device deployment.

推理Agent对话
ollama run lfm2.5-thinking
magistral

Community · 24b · unknown · 本地优先 · ollama-library · 1.4M pulls

Q4 ~14GB舒适 24GB

Magistral is a small, efficient reasoning model with 24B parameters.

推理Agent对话
ollama run magistral
minimax-m2.7

Community · ~7B · unknown · 本地优先 · ollama-library · 2.3M pulls

Q4 ~4GB舒适 8GB

MiniMax's M2-series model for coding, agentic workflows, and professional productivity.

代码推理Agent对话
ollama run minimax-m2.7
ministral-3

Community · ~7B · unknown · 本地优先 · ollama-library · 1.4M pulls

Q4 ~6GB舒适 12GB

The Ministral 3 family is designed for edge deployment, capable of running on a wide range of hardware.

视觉Agent对话
ollama run ministral-3
moondream

Community · ~7B · unknown · 本地优先 · ollama-library · 1.5M pulls

Q4 ~6GB舒适 12GB

moondream2 is a small vision language model designed to run efficiently on edge devices.

视觉对话
ollama run moondream
neural-chat

Community · ~7B · unknown · 本地优先 · ollama-library · 1.1M pulls

Q4 ~4GB舒适 8GB

A fine-tuned model based on Mistral with good coverage of domain and language.

对话
ollama run neural-chat
nous-hermes

Community · ~7B · unknown · 本地优先 · ollama-library · 1.2M pulls

Q4 ~4GB舒适 8GB

General use models based on Llama and Llama 2 from Nous Research.

对话
ollama run nous-hermes
nous-hermes2

Community · ~7B · unknown · 本地优先 · ollama-library · 1.1M pulls

Q4 ~4GB舒适 8GB

The powerful family of models by Nous Research that excels at scientific discussion and coding tasks.

代码对话
ollama run nous-hermes2
olmo2

Community · 7b/13b · unknown · 本地优先 · ollama-library · 3.7M pulls

Q4 ~4GB舒适 8GB

OLMo 2 is a new family of 7B and 13B models trained on up to 5T tokens. These models are on par with or better than equivalently sized fully open models, and competitive with open-weight models such as Llama 3.1 on English academic benchmarks.

对话
ollama run olmo2
openchat

Community · ~7B · unknown · 本地优先 · ollama-library · 1.2M pulls

Q4 ~4GB舒适 8GB

A family of open-source models trained on a wide variety of data, surpassing ChatGPT on various benchmarks. Updated to version 3.5-0106.

对话
ollama run openchat
openhermes

Community · 7b · unknown · 本地优先 · ollama-library · 1.1M pulls

Q4 ~4GB舒适 8GB

OpenHermes 2.5 is a 7B model fine-tuned by Teknium on Mistral with fully open datasets.

对话
ollama run openhermes
openthinker

Community · ~7B · unknown · 本地优先 · ollama-library · 1.1M pulls

Q4 ~4GB舒适 8GB

A fully open-source family of reasoning models built using a dataset derived by distilling DeepSeek-R1.

推理对话
ollama run openthinker
orca-mini

Community · ~7B · unknown · 本地优先 · ollama-library · 3.0M pulls

Q4 ~4GB舒适 8GB

A general-purpose model ranging from 3 billion parameters to 70 billion, suitable for entry-level hardware.

对话
ollama run orca-mini
smollm

Community · 1.7b · unknown · 本地优先 · ollama-library · 2.1M pulls

Q4 ~4GB舒适 8GB

🪐 A family of small models with 135M, 360M, and 1.7B parameters, trained on a new high-quality dataset.

对话
ollama run smollm
smollm2

Community · 1.7b · unknown · 本地优先 · ollama-library · 3.9M pulls

Q4 ~4GB舒适 8GB

SmolLM2 is a family of compact language models available in three size: 135M, 360M, and 1.7B parameters.

Agent对话
ollama run smollm2
sqlcoder

Community · ~7B · unknown · 本地优先 · ollama-library · 1.4M pulls

Q4 ~4GB舒适 8GB

SQLCoder is a code completion model fined-tuned on StarCoder for SQL generation tasks

代码对话
ollama run sqlcoder
stable-code

Community · 3b/7b · unknown · 本地优先 · ollama-library · 1.0M pulls

Q4 ~4GB舒适 8GB

Stable Code 3B is a coding model with instruct and code completion variants on par with models such as Code Llama 7B that are 2.5x larger.

代码对话
ollama run stable-code
stablelm2

Community · 1.6b/12b · unknown · 本地优先 · ollama-library · 1.0M pulls

Q4 ~8GB舒适 16GB

Stable LM 2 is a state-of-the-art 1.6B and 12B parameter language model trained on multilingual data in English, Spanish, German, Italian, French, Portuguese, and Dutch.

对话多语
ollama run stablelm2
starcoder

Community · ~7B · unknown · 本地优先 · ollama-library · 1.2M pulls

Q4 ~4GB舒适 8GB

StarCoder is a code generation model trained on 80+ programming languages.

代码对话
ollama run starcoder
starcoder2

Community · 3b/7b/15b · unknown · 本地优先 · ollama-library · 2.9M pulls

Q4 ~4GB舒适 8GB

StarCoder2 is the next generation of transparently trained open code LLMs that comes in three sizes: 3B, 7B and 15B parameters.

代码对话
ollama run starcoder2
vicuna

Community · ~7B · unknown · 本地优先 · ollama-library · 1.2M pulls

Q4 ~4GB舒适 8GB

General use chat model based on Llama and Llama 2 with 2K to 16K context sizes.

对话
ollama run vicuna
wizard-vicuna-uncensored

Community · 7b/13b/30b · unknown · 本地优先 · ollama-library · 1.2M pulls

Q4 ~4GB舒适 8GB

Wizard Vicuna Uncensored is a 7B, 13B, and 30B parameter model based on Llama 2 uncensored by Eric Hartford.

对话
ollama run wizard-vicuna-uncensored
wizardcoder

Community · ~7B · unknown · 本地优先 · ollama-library · 1.0M pulls

Q4 ~4GB舒适 8GB

State-of-the-art code generation model

代码对话
ollama run wizardcoder

与竞品对齐的能力

  • ApX / VRAM 计算器类:本站 /tools/local-deploy 覆盖显存适配、多卡 TP、产出与 vs API 费用。
  • Ollama Library / HF Hub:每条模型直链下载与 ollama run,不二次托管权重。
  • Model Picker(按 GPU 筛模型):列表「显存上限」筛选 + 计算器预填。
  • 价格/中转对照:有 OpenRouter id 的可回 /models · /official-api
Open-source models hub · HF / Ollama / GGUF deploy · GrokCode 倍率榜