Ollama Models

AI Models for Ollama

Every model in this directory is GGUF format, so it runs directly in Ollama. Here is the full catalog with ready-to-copy pull commands.

gemma-4-31B-it-GGUF

Google31B24 GB RAM

31B open-weight model from Google for local AI inference.

ollama pull gemma-4-31b-it:31b

Gemma-4-E4B-Uncensored-HauhauCS-Aggressive

HauhauCS4B8 GB RAM

4B open-weight model from HauhauCS for local AI inference.

ollama pull gemma-4-e4b-uncensored-hauhaucs-aggressive:4b

Qwopus3.6-27B-Coder-Compat-MTP-GGUF

Jackrong27B24 GB RAM

27B open-weight model from Jackrong for local AI inference.

ollama pull qwopus3.6-27b-coder-compat-mtp:27b

UI-TARS-1.5-7B-GGUF

mradermacher7B8 GB RAM

7B open-weight model from mradermacher for local AI inference.

ollama pull ui-tars-1.5-7b:7b

Huihui-DeepSeek-V4-Flash-abliterated-ds4-GGUF

huihui-aiUnknown8 GB RAM

Unknown open-weight model from huihui-ai for local AI inference.

ollama pull huihui-deepseek-v4-flash-abliterated-ds4:latest

gemma-4-12B-it-QAT-GGUF

Google12B16 GB RAM

12B open-weight model from Google for local AI inference.

ollama pull gemma-4-12b-it-qat:12b

Ternary-Bonsai-27B-gguf

prism-ml27B24 GB RAM

27B open-weight model from prism-ml for local AI inference.

ollama pull ternary-bonsai-27b:27b

Gemmable-4-12B-MTP-GGUF

Mia-AiLab12B16 GB RAM

12B open-weight model from Mia-AiLab for local AI inference.

ollama pull gemmable-4-12b-mtp:12b

Qwopus3.6-35B-A3B-Coder-MTP-GGUF

Jackrong35B48 GB RAM

35B open-weight model from Jackrong for local AI inference.

ollama pull qwopus3.6-35b-a3b-coder-mtp:35b

Qwen3.6-35B-A3B-MTP-GGUF

Alibaba35B48 GB RAM

35B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.6-35b-a3b-mtp:35b

gemma-4-12b-it-GGUF

Google12B16 GB RAM

12B open-weight model from Google for local AI inference.

ollama pull gemma-4-12b-it:12b

Qwen3-VL-30B-A3B-Instruct-GGUF

Qwen30B24 GB RAM

30B open-weight model from Qwen for local AI inference.

ollama pull qwen3-vl-30b-a3b:30b

PaddleOCR-VL-1.6-GGUF

PaddlePaddleUnknown8 GB RAM

Unknown open-weight model from PaddlePaddle for local AI inference.

ollama pull paddleocr-vl-1.6:latest

models-moved

ggml-orgUnknown8 GB RAM

Unknown open-weight model from ggml-org for local AI inference.

ollama pull models-moved:latest

cohere-transcribe-03-2026-gguf

handy-computerUnknown8 GB RAM

Unknown open-weight model from handy-computer for local AI inference.

ollama pull cohere-transcribe-03-2026:latest

Qwen-AgentWorld-35B-A3B-GGUF

Alibaba35B48 GB RAM

35B open-weight model from Alibaba for local AI inference.

ollama pull qwen-agentworld-35b-a3b:35b

vntl-llama3-8b-v2-gguf

lmg-anon8B16 GB RAM

8B open-weight model from lmg-anon for local AI inference.

ollama pull vntl-llama3-8b-v2:8b

gemma-4-E4B-it-GGUF

Google4B8 GB RAM

4B open-weight model from Google for local AI inference.

ollama pull gemma-4-e4b-it:4b

Qwen3.6-27B-GGUF

Alibaba27B24 GB RAM

27B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.6-27b:27b

Qwen3.6-35B-A3B-GGUF

Alibaba35B48 GB RAM

35B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.6-35b-a3b:35b

Hy3-GGUF

vcruz305Unknown8 GB RAM

Unknown open-weight model from vcruz305 for local AI inference.

ollama pull hy3:latest

HyperCLOVAX-SEED-Text-Instruct-1.5B-Q4_K_M-GGUF

rippertnt1.5B4 GB RAM

1.5B open-weight model from rippertnt for local AI inference.

ollama pull hyperclovax-seed-text-instruct-1.5b-q4-k-m:1.5b

Qwen3.5-9B-GGUF

Alibaba9B16 GB RAM

9B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.5-9b:9b

Qwen3.5-4B-GGUF

Alibaba4B8 GB RAM

4B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.5-4b:4b

parakeet-unified-en-0.6b-gguf

handy-computer0.6B2 GB RAM

0.6B open-weight model from handy-computer for local AI inference.

ollama pull parakeet-unified-en-0.6b:0.6b

nemotron-3.5-asr-streaming-0.6b-gguf

handy-computer0.6B2 GB RAM

0.6B open-weight model from handy-computer for local AI inference.

ollama pull nemotron-3.5-asr-streaming-0.6b:0.6b

gemma-4-26B-A4B-it-GGUF

Google26B24 GB RAM

26B open-weight model from Google for local AI inference.

ollama pull gemma-4-26b-a4b-it:26b

Bonsai-27B-gguf

prism-ml27B24 GB RAM

27B open-weight model from prism-ml for local AI inference.

ollama pull bonsai-27b:27b

Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive

HauhauCS35B48 GB RAM

35B open-weight model from HauhauCS for local AI inference.

ollama pull qwen3.6-35b-a3b-uncensored-hauhaucs-aggressive:35b

Qwythos-9B-Claude-Mythos-5-1M-GGUF

empero-ai9B16 GB RAM

9B open-weight model from empero-ai for local AI inference.

ollama pull qwythos-9b-claude-mythos-5-1m:9b

Ornith-1.0-35B-GGUF

deepreinforce-ai35B48 GB RAM

35B open-weight model from deepreinforce-ai for local AI inference.

ollama pull ornith-1.0-35b:35b

Qwen3.6-27B-MTP-GGUF

Alibaba27B24 GB RAM

27B open-weight model from Alibaba for local AI inference.

ollama pull qwen3.6-27b-mtp:27b

mxbai-embed-large-v1

mixedbread-aiUnknown8 GB RAM

Unknown open-weight model from mixedbread-ai for local AI inference.

ollama pull mxbai-embed-large-v1:latest

embeddinggemma-300M-GGUF

ggml-orgUnknown8 GB RAM

Unknown open-weight model from ggml-org for local AI inference.

ollama pull embeddinggemma-300m:latest

deepseek-v4-gguf

antirezUnknown8 GB RAM

Unknown open-weight model from antirez for local AI inference.

ollama pull deepseek-v4:latest

Ornith-1.0-9B-GGUF

deepreinforce-ai9B16 GB RAM

9B open-weight model from deepreinforce-ai for local AI inference.

ollama pull ornith-1.0-9b:9b

Hello, Nice to meet you! 👋

Subscribe our newsletter to get the latest AI news.