Open Models · Local Deploy

全球开源模型频道

策展 + Ollama Library 全量 + HF GGUF 热门 合并目录:参数/显存、许可、 Hugging Face / GGUF / Ollama / ModelScope 下载,并跳转 本地算力计算器 。HF 全库十万+——本频道做可跑清单、深链与算力账,不镜像权重。 同步于 2026-08-03T07:33:50Z

合并总量325
策展精选42
同步条数303
本地优先318
含 Ollama238
含 HF/GGUF192

全库入口(全球)

要「全部开源」请用下列权威源检索;本站策展负责选型 + 部署路径 + 算力账

按家族

我的显卡一键筛:

匹配 222 · 展示 40 · 点「设为 A/B」后底部一键对比 · 本站不托管权重

llama3.2

Meta · 1b/3b · unknown · 本地优先 · ollama-library · 79M pulls

Q4 ~4GB舒适 8GB

Meta's Llama 3.2 goes small with 1B and 3B models.

Agent对话
ollama run llama3.2
gemma4

Google · ~7B · unknown · 本地优先 · ollama-library · 20M pulls

Q4 ~6GB舒适 12GB

Gemma 4 models are designed to deliver frontier-level performance at each size. They are well-suited for reasoning, agentic workflows, coding, and multimodal understanding.

视觉代码推理Agent对话
ollama run gemma4
llama3

Meta · ~7B · unknown · 本地优先 · ollama-library · 25M pulls

Q4 ~4GB舒适 8GB

Meta Llama 3: The most capable openly available LLM to date

对话
ollama run llama3
qwen2.5-coder

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 20M pulls

Q4 ~4GB舒适 8GB

The latest series of Code-Specific Qwen models, with significant improvements in code generation, code reasoning, and code fixing.

代码推理Agent对话
ollama run qwen2.5-coder
codellama

Meta · ~7B · unknown · 本地优先 · ollama-library · 5.9M pulls

Q4 ~4GB舒适 8GB

A large language model that can use text prompts to generate and discuss code.

代码对话
ollama run codellama
gemma

Google · ~7B · unknown · 本地优先 · ollama-library · 7.3M pulls

Q4 ~4GB舒适 8GB

Gemma is a family of lightweight, state-of-the-art open models built by Google DeepMind. Updated to version 1.1

对话
ollama run gemma
glm-ocr

Zhipu · ~7B · unknown · 本地优先 · ollama-library · 6.2M pulls

Q4 ~6GB舒适 12GB

GLM-OCR is a multimodal OCR model for complex document understanding, built on the GLM-V encoder–decoder architecture.

视觉OCR代码Agent
ollama run glm-ocr
gpt-oss

Community · ~7B · unknown · 本地优先 · ollama-library · 11M pulls

Q4 ~4GB舒适 8GB

OpenAI’s open-weight models designed for powerful reasoning, agentic tasks, and versatile developer use cases.

推理Agent对话
ollama run gpt-oss
llama2

Meta · 7b/70b · unknown · 本地优先 · ollama-library · 7.3M pulls

Q4 ~4GB舒适 8GB

Llama 2 is a collection of foundation language models ranging from 7B to 70B parameters.

对话
ollama run llama2
minicpm-v

Community · ~7B · unknown · 本地优先 · ollama-library · 5.4M pulls

Q4 ~6GB舒适 12GB

A series of multimodal LLMs (MLLMs) designed for vision-language understanding.

视觉对话
ollama run minicpm-v
mistral-nemo

Mistral AI · 12b · unknown · 本地优先 · ollama-library · 5.6M pulls

Q4 ~8GB舒适 16GB

A state-of-the-art 12B model with 128k context length, built by Mistral AI in collaboration with NVIDIA.

Agent对话
ollama run mistral-nemo
mxbai-embed-large

Community · ~1B · unknown · 本地优先 · ollama-library · 13M pulls

Q4 ~0.8GB舒适 2GB

State-of-the-art large embedding model from mixedbread.ai

向量
ollama run mxbai-embed-large
phi3

Microsoft · 3b/14b · unknown · 本地优先 · ollama-library · 18M pulls

Q4 ~4GB舒适 8GB

Phi-3 is a family of lightweight 3B (Mini) and 14B (Medium) state-of-the-art open models by Microsoft.

对话
ollama run phi3
qwen

Alibaba · 0.5b/110b · unknown · 本地优先 · ollama-library · 7.5M pulls

Q4 ~38.5GB舒适 49GB

Qwen 1.5 is a series of large language models by Alibaba Cloud spanning from 0.5B to 110B parameters

对话
ollama run qwen
qwen2

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 6.1M pulls

Q4 ~4GB舒适 8GB

Qwen2 is a new series of large language models from Alibaba group

Agent对话
ollama run qwen2
qwen3-vl

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 5.0M pulls

Q4 ~6GB舒适 12GB

The most powerful vision-language model in the Qwen model family to date.

视觉推理Agent对话
ollama run qwen3-vl
qwen3.5

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 17M pulls

Q4 ~6GB舒适 12GB

Qwen 3.5 is a family of open-source multimodal models that delivers exceptional utility and performance.

视觉推理Agent对话
ollama run qwen3.5
qwen3.6

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 5.1M pulls

Q4 ~6GB舒适 12GB

Qwen3.6 delivers substantial upgrades in agentic coding and thinking preservation than previous Qwen models.

视觉代码推理Agent对话
ollama run qwen3.6
tinyllama

Meta · 1.1b · unknown · 本地优先 · ollama-library · 5.3M pulls

Q4 ~4GB舒适 8GB

The TinyLlama project is an open endeavor to train a compact 1.1B Llama model on 3 trillion tokens.

对话
ollama run tinyllama
all-minilm

Community · ~1B · unknown · 本地优先 · ollama-library · 3.3M pulls

Q4 ~0.8GB舒适 2GB

Embedding models on very large sentence level datasets.

向量
ollama run all-minilm
aya

Community · ~7B · unknown · 本地优先 · ollama-library · 1.1M pulls

Q4 ~4GB舒适 8GB

Aya 23, released by Cohere, is a new family of state-of-the-art, multilingual models that support 23 languages.

对话多语
ollama run aya
codegemma

Google · ~7B · unknown · 本地优先 · ollama-library · 3.1M pulls

Q4 ~4GB舒适 8GB

CodeGemma is a collection of powerful, lightweight models that can perform a variety of coding tasks like fill-in-the-middle code completion, code generation, natural language understanding, mathematical reasoning, and instruction following.

代码推理数学对话
ollama run codegemma
codeqwen

Alibaba · ~7B · unknown · 本地优先 · ollama-library · 1.1M pulls

Q4 ~4GB舒适 8GB

CodeQwen1.5 is a large language model pretrained on a large amount of code data.

代码对话
ollama run codeqwen
codestral

Mistral AI · ~7B · unknown · 本地优先 · ollama-library · 1.3M pulls

Q4 ~4GB舒适 8GB

Codestral is Mistral AI’s first-ever code model designed for code generation tasks.

代码对话
ollama run codestral
cogito

Community · ~7B · unknown · 本地优先 · ollama-library · 2.1M pulls

Q4 ~4GB舒适 8GB

Cogito v1 Preview is a family of hybrid reasoning models by Deep Cogito that outperform the best available open models of the same size, including counterparts from LLaMA, DeepSeek, and Qwen across most standard benchmarks.

推理Agent对话
ollama run cogito
command-r

Community · ~7B · unknown · 本地优先 · ollama-library · 1.5M pulls

Q4 ~4GB舒适 8GB

Command R is a Large Language Model optimized for conversational interaction and long context tasks.

Agent对话
ollama run command-r
deepscaler

Community · 1.5b · unknown · 本地优先 · ollama-library · 1.3M pulls

Q4 ~4GB舒适 8GB

A fine-tuned version of Deepseek-R1-Distilled-Qwen-1.5B that surpasses the performance of OpenAI’s o1-preview with just 1.5B parameters on popular math evaluations.

推理数学对话
ollama run deepscaler
deepseek-coder

DeepSeek · ~7B · unknown · 本地优先 · ollama-library · 4.5M pulls

Q4 ~4GB舒适 8GB

DeepSeek Coder is a capable coding model trained on two trillion code and natural language tokens.

代码对话
ollama run deepseek-coder
dolphin-llama3

Meta · 8b/70b · unknown · 本地优先 · ollama-library · 2.0M pulls

Q4 ~4.4GB舒适 8.8GB

Dolphin 2.9 is a new model with 8B and 70B sizes by Eric Hartford based on Llama 3 that has a variety of instruction, conversational, and coding skills.

代码对话
ollama run dolphin-llama3
dolphin-mistral

Mistral AI · ~7B · unknown · 本地优先 · ollama-library · 1.6M pulls

Q4 ~4GB舒适 8GB

The uncensored Dolphin model based on Mistral that excels at coding tasks. Updated to version 2.8.

代码对话
ollama run dolphin-mistral
dolphin-mixtral

Mistral AI · 8x7b/8x22b · unknown · 本地优先 · ollama-library · 1.8M pulls

Q4 ~38.5GB舒适 49GB

Uncensored, 8x7b and 8x22b fine-tuned models based on the Mixtral mixture of experts models that excels at coding tasks. Created by Eric Hartford.

代码MoE对话
ollama run dolphin-mixtral
dolphin-phi

Microsoft · 2.7b · unknown · 本地优先 · ollama-library · 1.6M pulls

Q4 ~4GB舒适 8GB

2.7B uncensored Dolphin model by Eric Hartford, based on the Phi language model by Microsoft Research.

对话
ollama run dolphin-phi

与竞品对齐的能力

  • ApX / VRAM 计算器类:本站 /tools/local-deploy 覆盖显存适配、多卡 TP、产出与 vs API 费用。
  • Ollama Library / HF Hub:每条模型直链下载与 ollama run,不二次托管权重。
  • Model Picker(按 GPU 筛模型):列表「显存上限」筛选 + 计算器预填。
  • 价格/中转对照:有 OpenRouter id 的可回 /models · /official-api
开源模型部署频道 · HF / Ollama / GGUF · GrokCode 倍率榜