Ollama Models

内容

AI Models for Ollama

Every model in this directory is GGUF format, so it runs directly in Ollama. Here is the full catalog with ready-to-copy pull commands.

Higgs-Audio-v3-Studio

drbaph未知8 GB的RAM

Unknown open-weight model from drbaph for local AI inference.

ollama pull higgs-audio-v3-studio:latest

DeepSeek-V4-Flash-GGUF

DeepSeek未知8 GB的RAM

Unknown open-weight model from DeepSeek for local AI inference.

ollama pull deepseek-v4-flash:latest

Qwen3.5-2B-GGUF

如阿里巴巴2B4 GB的RAM

2B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.5-2b:2b

Qwen3.5-4B-Uncensored-HauhauCS-Aggressive

HauhauCS4B8 GB的RAM

4B open-weight model from HauhauCS for local AI inference.

ollama pull qwen3.5-4b-uncensored-hauhaucs-aggressive:4b

Qwen3.5-35B-A3B-GGUF

如阿里巴巴35B48 GB的RAM

35B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.5-35b-a3b:35b

Llama-3.2-3B-Instruct-GGUF

3B4 GB的RAM

3B open-weight model from Meta for local AI inference.

ollama pull llama-3.2-3b:3b

gemma-3-1b-it-GGUF

ggml-org1B2 GB的RAM

1B open-weight model from ggml-org for local AI inference.

ollama pull gemma-3-1b-it:1b

Qwen3.5-122B-A10B-MTP-GGUF

如阿里巴巴122B80 GB的RAM

122B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.5-122b-a10b-mtp:122b

FLUX.2-klein-4B-GGUF

不懒惰4B8 GB的RAM

4B open-weight model from unsloth for local AI inference.

ollama pull flux.2-klein-4b:4b

Qwen2.5-7B-Instruct-GGUF

如阿里巴巴7B8 GB的RAM

7B open-weight model from Alibaba for local AI inference.

ollama pull qwen2.5-7b:7b

Qwen3-30B-A3B-GGUF

MaziyarPanahi30B24 GB的RAM

30B open-weight model from MaziyarPanahi for local AI inference.

ollama pull qwen3-30b-a3b:30b

Qwen3-32B-GGUF

MaziyarPanahi32B24 GB的RAM

32B open-weight model from MaziyarPanahi for local AI inference.

ollama pull qwen3-32b:32b

Qwen3-1.7B-GGUF

MaziyarPanahi1.7B4 GB的RAM

1.7B open-weight model from MaziyarPanahi for local AI inference.

ollama pull qwen3-1.7b:1.7b

Kimi-K3-GGUF

不懒惰未知8 GB的RAM

Unknown open-weight model from unsloth for local AI inference.

ollama pull kimi-k3:latest

Qwen2.5-0.5B-Instruct-GGUF

wen文0.5B2 GB的RAM

0.5B open-weight model from Qwen for local AI inference.

ollama pull qwen2.5-0.5b:0.5b

Qwen3-14B-GGUF

MaziyarPanahi14B24 GB的RAM

14B open-weight model from MaziyarPanahi for local AI inference.

ollama pull qwen3-14b:14b

Qwen3-0.6B-GGUF

MaziyarPanahi0.6B2 GB的RAM

0.6B open-weight model from MaziyarPanahi for local AI inference.

ollama pull qwen3-0.6b:0.6b

Laguna-S-2.1-GGUF

不懒惰未知8 GB的RAM

Unknown open-weight model from unsloth for local AI inference.

ollama pull laguna-s-2.1:latest

Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V7-GGUF

LuffyTheFox35B48 GB的RAM

35B open-weight model from LuffyTheFox for local AI inference.

ollama pull qwen3.6-35b-a3b-uncensored-genesis-hermes-v7:35b

gemma-gguf

rahul7star未知8 GB的RAM

Unknown open-weight model from rahul7star for local AI inference.

ollama pull gemma:latest

Qwen2.5-VL-7B-Instruct-GGUF

如阿里巴巴7B8 GB的RAM

7B open-weight model from Alibaba for local AI inference.

ollama pull qwen2.5-vl-7b:7b

Qwen3-235B-A22B-GGUF

如阿里巴巴235B80 GB的RAM

235B open-weight model from Alibaba for local AI inference.

ollama pull qwen3-235b-a22b:235b

Voxtral-Small-24B-2507-gguf

handy-computer24B24 GB的RAM

24B open-weight model from handy-computer for local AI inference.

ollama pull voxtral-small-24b-2507:24b

Sugoi-14B-Ultra-GGUF

sugoitoolkit14B24 GB的RAM

14B open-weight model from sugoitoolkit for local AI inference.

ollama pull sugoi-14b-ultra:14b

Qwen3-ASR-1.7B-gguf

handy-computer1.7B4 GB的RAM

1.7B open-weight model from handy-computer for local AI inference.

ollama pull qwen3-asr-1.7b:1.7b

Qwen3-30B-A3B-Instruct-2507-GGUF

如阿里巴巴30B24 GB的RAM

30B open-weight model from Alibaba for local AI inference.

ollama pull qwen3-30b-a3b-instruct-2507:30b

Qwen2.5-3B-Instruct-GGUF

wen文3B4 GB的RAM

3B open-weight model from Qwen for local AI inference.

ollama pull qwen2.5-3b:3b

gemma-4-12b-heretic-abliterated-GGUF

culturerevolt12B16 GB的RAM

12B open-weight model from culturerevolt for local AI inference.

ollama pull gemma-4-12b-heretic-abliterated:12b

Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive

HauhauCS122B80 GB的RAM

122B open-weight model from HauhauCS for local AI inference.

ollama pull qwen3.5-122b-a10b-uncensored-hauhaucs-aggressive:122b

Qwen2.5-Coder-7B-Instruct-GGUF

wen文7B8 GB的RAM

7B open-weight model from Qwen for local AI inference.

ollama pull qwen2.5-coder-7b:7b

Qwen-Image-Edit-2511-GGUF

如阿里巴巴未知8 GB的RAM

Unknown open-weight model from Alibaba for local AI inference.

ollama pull qwen-image-edit-2511:latest

Qwen3-Coder-Next-GGUF

如阿里巴巴未知8 GB的RAM

Unknown open-weight model from Alibaba for local AI inference.

ollama pull qwen3-coder-next:latest

Qwen2.5-1.5B-Instruct-GGUF

wen文1.5B4 GB的RAM

1.5B open-weight model from Qwen for local AI inference.

ollama pull qwen2.5-1.5b:1.5b

Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-GGUF

LuffyTheFox35B48 GB的RAM

35B open-weight model from LuffyTheFox for local AI inference.

ollama pull qwen3.6-35b-a3b-uncensored-genesis-hermes-v6:35b

paw-programs

programasweights未知8 GB的RAM

Unknown open-weight model from programasweights for local AI inference.

ollama pull paw-programs:latest

Qwen2.5-Coder-32B-Instruct-GGUF

wen文32B24 GB的RAM

32B open-weight model from Qwen for local AI inference.

ollama pull qwen2.5-coder-32b:32b

gemma-4-31B-it-qat-GGUF

Google31B24 GB的RAM

31B open-weight model from Google for local AI inference.

ollama pull gemma-4-31b-it-qat:31b

gemma-4-26B-A4B-it-ultra-uncensored-heretic-i1-GGUF

Google26B24 GB的RAM

26B open-weight model from Google for local AI inference.

ollama pull gemma-4-26b-a4b-it-ultra-uncensored-heretic-i1:26b

gemma-4-12B-coder-fable5-composer2.5-v1-GGUF

yuxinlu112B16 GB的RAM

12B open-weight model from yuxinlu1 for local AI inference.

ollama pull gemma-4-12b-coder-fable5-composer2.5-v1:12b

inkling-GGUF

不懒惰未知8 GB的RAM

Unknown open-weight model from unsloth for local AI inference.

ollama pull inkling:latest

granite-4.1-3b-GGUF

ibm-granite3B4 GB的RAM

3B open-weight model from ibm-granite for local AI inference.

ollama pull granite-4.1-3b:3b

Sugoi-32B-Ultra-GGUF

sugoitoolkit32B24 GB的RAM

32B open-weight model from sugoitoolkit for local AI inference.

ollama pull sugoi-32b-ultra:32b

Qwen3-VL-4B-Instruct-GGUF

如阿里巴巴4B8 GB的RAM

4B open-weight model from Alibaba for local AI inference.

ollama pull qwen3-vl-4b:4b

madlad400-3b-mt

谷歌3B4 GB的RAM

3B open-weight model from google for local AI inference.

ollama pull madlad400-3b-mt:3b

MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking-GGUF

GnLOLot1B2 GB的RAM

1B open-weight model from GnLOLot for local AI inference.

ollama pull minicpm5-1b-claude-opus-fable5-v2-thinking:1b

svara-tts-v1

kenpath未知8 GB的RAM

Unknown open-weight model from kenpath for local AI inference.

ollama pull svara-tts-v1:latest

Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-NEO-CODE-Di-IMatrix-MAX-GGUF

DavidAU40B48 GB的RAM

40B open-weight model from DavidAU for local AI inference.

ollama pull qwen3.6-40b-claude-4.6-opus-deckard-heretic-uncensored-thinking-neo-code-di-imatrix-max:40b

Holo-3.1-35B-A3B-GGUF

Hcompany35B48 GB的RAM

35B open-weight model from Hcompany for local AI inference.

ollama pull holo-3.1-35b-a3b:35b

gemma-4-12B-it-qat-q4_0-gguf

谷歌12B16 GB的RAM

12B open-weight model from google for local AI inference.

ollama pull gemma-4-12b-it-qat-q4-0:12b

Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF

DavidAU9B16 GB的RAM

9B open-weight model from DavidAU for local AI inference.

ollama pull qwen3.5-9b-the-defiant-fable-uncensored-heretic-neo-imatrix-max-mtp:9b

LTX-2.3-GGUF

不懒惰未知8 GB的RAM

Unknown open-weight model from unsloth for local AI inference.

ollama pull ltx-2.3:latest

gemma-4-E4B-it-qat-GGUF

Google4B8 GB的RAM

4B open-weight model from Google for local AI inference.

ollama pull gemma-4-e4b-it-qat:4b

whisper-large-v3-turbo-gguf

handy-computer未知8 GB的RAM

Unknown open-weight model from handy-computer for local AI inference.

ollama pull whisper-large-v3-turbo:latest

LTX-2

DeepBeepMeep未知8 GB的RAM

Unknown open-weight model from DeepBeepMeep for local AI inference.

ollama pull ltx-2:latest

Qwen3-4B-GGUF

wen文4B8 GB的RAM

4B open-weight model from Qwen for local AI inference.

ollama pull qwen3-4b:4b

Qwen3.5-0.8B-GGUF

如阿里巴巴0.8B2 GB的RAM

0.8B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.5-0.8b:0.8b

Meta-Llama-3.1-8B-Instruct-GGUF

8B16 GB的RAM

8B open-weight model from Meta for local AI inference.

ollama pull meta-llama-3.1-8b:8b

LFM2.5-1.2B-Instruct-GGUF

LiquidAI1.2B4 GB的RAM

1.2B open-weight model from LiquidAI for local AI inference.

ollama pull lfm2.5-1.2b:1.2b

GLM-5.2-GGUF

不懒惰未知8 GB的RAM

Unknown open-weight model from unsloth for local AI inference.

ollama pull glm-5.2:latest

canary-180m-flash-gguf

handy-computer未知8 GB的RAM

Unknown open-weight model from handy-computer for local AI inference.

ollama pull canary-180m-flash:latest

whisper-large-v3-gguf

handy-computer未知8 GB的RAM

Unknown open-weight model from handy-computer for local AI inference.

ollama pull whisper-large-v3:latest

Qwen3-8B-GGUF

如阿里巴巴8B16 GB的RAM

8B open-weight model from Alibaba for local AI inference.

ollama pull qwen3-8b:8b

Jan-v3.5-4B-gguf

janhq4B8 GB的RAM

4B open-weight model from janhq for local AI inference.

ollama pull jan-v3.5-4b:4b

Kimi-K2.7-Code-GGUF

不懒惰未知8 GB的RAM

Unknown open-weight model from unsloth for local AI inference.

ollama pull kimi-k2.7-code:latest

MiniCPM5-1B-Claude-Opus-Fable5-Thinking-GGUF

GnLOLot1B2 GB的RAM

1B open-weight model from GnLOLot for local AI inference.

ollama pull minicpm5-1b-claude-opus-fable5-thinking:1b

Llama-3.2-1B-Instruct-Q8_0-GGUF

hugging-quants1B2 GB的RAM

1B open-weight model from hugging-quants for local AI inference.

ollama pull llama-3.2-1b-instruct-q8-0:1b

Bielik-11B-v3.0-Instruct-awq

隐形狗绳11B16 GB的RAM

11B open-weight model from speakleash for local AI inference.

ollama pull bielik-11b-v3.0-instruct-awq:11b

Wan2.2-I2V-A14B-GGUF

QuantStack14B24 GB的RAM

14B open-weight model from QuantStack for local AI inference.

ollama pull wan2.2-i2v-a14b:14b

Gemma-4-31B-it-abliterated

paperscarecrow31B24 GB的RAM

31B open-weight model from paperscarecrow for local AI inference.

ollama pull gemma-4-31b-it-abliterated:31b

Qwen3.6-27B-Uncensored-HauhauCS-Aggressive

HauhauCS27B24 GB的RAM

27B open-weight model from HauhauCS for local AI inference.

ollama pull qwen3.6-27b-uncensored-hauhaucs-aggressive:27b

ThinkingCap-Qwen3.6-27B-GGUF

bottlecapai27B24 GB的RAM

27B open-weight model from bottlecapai for local AI inference.

ollama pull thinkingcap-qwen3.6-27b:27b

万新

DeepBeepMeep未知8 GB的RAM

Unknown open-weight model from DeepBeepMeep for local AI inference.

ollama pull wan2.1:latest

Voxtral-Mini-4B-Realtime-2602-gguf

handy-computer4B8 GB的RAM

4B open-weight model from handy-computer for local AI inference.

ollama pull voxtral-mini-4b-realtime-2602:4b

Qwen_Qwen3.6-35B-A3B-GGUF

如阿里巴巴35B48 GB的RAM

35B open-weight model from Alibaba for local AI inference.

ollama pull qwen-qwen3.6-35b-a3b:35b

Qwythos-9B-v2-GGUF

empero-ai9B16 GB的RAM

9B open-weight model from empero-ai for local AI inference.

ollama pull qwythos-9b-v2:9b

gemma-4-12B-agentic-fable5-composer2.5-v2-3.5x-tau2-GGUF

yuxinlu112B16 GB的RAM

12B open-weight model from yuxinlu1 for local AI inference.

ollama pull gemma-4-12b-agentic-fable5-composer2.5-v2-3.5x-tau2:12b

Sulphur-2-base

SulphurAI未知8 GB的RAM

Unknown open-weight model from SulphurAI for local AI inference.

ollama pull sulphur-2-base:latest

whisper-medium-gguf

handy-computer未知8 GB的RAM

Unknown open-weight model from handy-computer for local AI inference.

ollama pull whisper-medium:latest

gemma-4-E2B-it-GGUF

Google2B4 GB的RAM

2B open-weight model from Google for local AI inference.

ollama pull gemma-4-e2b-it:2b

parakeet-tdt-0.6b-v3-gguf

handy-computer0.6B2 GB的RAM

0.6B open-weight model from handy-computer for local AI inference.

ollama pull parakeet-tdt-0.6b-v3:0.6b

Qwen3.5-9B-Uncensored-HauhauCS-Aggressive

HauhauCS9B16 GB的RAM

9B open-weight model from HauhauCS for local AI inference.

ollama pull qwen3.5-9b-uncensored-hauhaucs-aggressive:9b

gemma-4-26B-A4B-it-qat-GGUF

Google26B24 GB的RAM

26B open-weight model from Google for local AI inference.

ollama pull gemma-4-26b-a4b-it-qat:26b

gpt-oss-20b-GGUF

不懒惰20B24 GB的RAM

20B open-weight model from unsloth for local AI inference.

ollama pull gpt-oss-20b:20b

Qwen3-VL-8B-Instruct-abliterated-GGUF

如阿里巴巴8B16 GB的RAM

8B open-weight model from Alibaba for local AI inference.

ollama pull qwen3-vl-8b-instruct-abliterated:8b

Flux2-Klein-9B-True-V2

wikeeyang9B16 GB的RAM

9B open-weight model from wikeeyang for local AI inference.

ollama pull flux2-klein-9b-true-v2:9b

Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF

DavidAU27B24 GB的RAM

27B open-weight model from DavidAU for local AI inference.

ollama pull qwen3.6-27b-fable-fusion-711-uncensored-heretic-nm-dau-neo-max-mtp:27b

Qwen3-Coder-30B-A3B-Instruct-GGUF

如阿里巴巴30B24 GB的RAM

30B open-weight model from Alibaba for local AI inference.

ollama pull qwen3-coder-30b-a3b:30b

gemma-4-31B-it-GGUF

Google31B24 GB的RAM

31B open-weight model from Google for local AI inference.

ollama pull gemma-4-31b-it:31b

Gemma-4-E4B-Uncensored-HauhauCS-Aggressive

HauhauCS4B8 GB的RAM

4B open-weight model from HauhauCS for local AI inference.

ollama pull gemma-4-e4b-uncensored-hauhaucs-aggressive:4b

Qwopus3.6​​-27B-Coder-Compat-MTP-GGUF

Jackrong27B24 GB的RAM

27B open-weight model from Jackrong for local AI inference.

ollama pull qwopus3.6-27b-coder-compat-mtp:27b

UI-TARS-1.5-7B-GGUF

mradermacher7B8 GB的RAM

7B open-weight model from mradermacher for local AI inference.

ollama pull ui-tars-1.5-7b:7b

Huihui-DeepSeek-V4-Flash-abliterated-ds4-GGUF

huihui-ai未知8 GB的RAM

Unknown open-weight model from huihui-ai for local AI inference.

ollama pull huihui-deepseek-v4-flash-abliterated-ds4:latest

gemma-4-12B-it-QAT-GGUF

Google12B16 GB的RAM

12B open-weight model from Google for local AI inference.

ollama pull gemma-4-12b-it-qat:12b

Ternary-Bonsai-27B-gguf

prism-ml27B24 GB的RAM

27B open-weight model from prism-ml for local AI inference.

ollama pull ternary-bonsai-27b:27b

Gemmable-4-12B-MTP-GGUF

Mia-AiLab12B16 GB的RAM

12B open-weight model from Mia-AiLab for local AI inference.

ollama pull gemmable-4-12b-mtp:12b

Qwopus3.6​​-35B-A3B-Coder-MTP-GGUF

Jackrong35B48 GB的RAM

35B open-weight model from Jackrong for local AI inference.

ollama pull qwopus3.6-35b-a3b-coder-mtp:35b

Qwen3.6-35B-A3B-MTP-GGUF

如阿里巴巴35B48 GB的RAM

35B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.6-35b-a3b-mtp:35b

gemma-4-12b-it-GGUF

Google12B16 GB的RAM

12B open-weight model from Google for local AI inference.

ollama pull gemma-4-12b-it:12b

Qwen3-VL-30B-A3B-Instruct-GGUF

wen文30B24 GB的RAM

30B open-weight model from Qwen for local AI inference.

ollama pull qwen3-vl-30b-a3b:30b

models-moved

ggml-org未知8 GB的RAM

Unknown open-weight model from ggml-org for local AI inference.

ollama pull models-moved:latest

cohere-transcribe-03-2026-gguf

handy-computer未知8 GB的RAM

Unknown open-weight model from handy-computer for local AI inference.

ollama pull cohere-transcribe-03-2026:latest

Qwen-AgentWorld-35B-A3B-GGUF

如阿里巴巴35B48 GB的RAM

35B open-weight model from Alibaba for local AI inference.

ollama pull qwen-agentworld-35b-a3b:35b

vntl-llama3-8b-v2-gguf

lmg-anon8B16 GB的RAM

8B open-weight model from lmg-anon for local AI inference.

ollama pull vntl-llama3-8b-v2:8b

gemma-4-E4B-it-GGUF

Google4B8 GB的RAM

4B open-weight model from Google for local AI inference.

ollama pull gemma-4-e4b-it:4b

Qwen3.6-27B-GGUF

如阿里巴巴27B24 GB的RAM

27B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.6-27b:27b

Qwen3.6-35B-A3B-GGUF

如阿里巴巴35B48 GB的RAM

35B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.6-35b-a3b:35b

Hy3-GGUF

vcruz305未知8 GB的RAM

Unknown open-weight model from vcruz305 for local AI inference.

ollama pull hy3:latest

HyperCLOVAX-SEED-Text-Instruct-1.5B-Q4_K_M-GGUF

rippertnt1.5B4 GB的RAM

1.5B open-weight model from rippertnt for local AI inference.

ollama pull hyperclovax-seed-text-instruct-1.5b-q4-k-m:1.5b

Qwen3.5-9B-GGUF

如阿里巴巴9B16 GB的RAM

9B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.5-9b:9b

Qwen3.5-4B-GGUF

如阿里巴巴4B8 GB的RAM

4B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.5-4b:4b

parakeet-unified-en-0.6b-gguf

handy-computer0.6B2 GB的RAM

0.6B open-weight model from handy-computer for local AI inference.

ollama pull parakeet-unified-en-0.6b:0.6b

nemotron-3.5-asr-streaming-0.6b-gguf

handy-computer0.6B2 GB的RAM

0.6B open-weight model from handy-computer for local AI inference.

ollama pull nemotron-3.5-asr-streaming-0.6b:0.6b

gemma-4-26B-A4B-it-GGUF

Google26B24 GB的RAM

26B open-weight model from Google for local AI inference.

ollama pull gemma-4-26b-a4b-it:26b

Bonsai-27B-gguf

prism-ml27B24 GB的RAM

27B open-weight model from prism-ml for local AI inference.

ollama pull bonsai-27b:27b

Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive

HauhauCS35B48 GB的RAM

35B open-weight model from HauhauCS for local AI inference.

ollama pull qwen3.6-35b-a3b-uncensored-hauhaucs-aggressive:35b

Qwythos-9B-Claude-Mythos-5-1M-GGUF

empero-ai9B16 GB的RAM

9B open-weight model from empero-ai for local AI inference.

ollama pull qwythos-9b-claude-mythos-5-1m:9b

Ornith-1.0-35B-GGUF

deepreinforce-ai35B48 GB的RAM

35B open-weight model from deepreinforce-ai for local AI inference.

ollama pull ornith-1.0-35b:35b

Qwen3.6-27B-MTP-GGUF

如阿里巴巴27B24 GB的RAM

27B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.6-27b-mtp:27b

deepseek-v4-gguf

antirez未知8 GB的RAM

Unknown open-weight model from antirez for local AI inference.

ollama pull deepseek-v4:latest

Ornith-1.0-9B-GGUF

deepreinforce-ai9B16 GB的RAM

9B open-weight model from deepreinforce-ai for local AI inference.

ollama pull ornith-1.0-9b:9b

Hello, Nice to meet you! 👋

Subscribe our newsletter to get the latest AI news.

Malcare WordPress 安全