Models

Sugoi-32B-Ultra-GGUF

sugoitoolkit32B24 GB RAM

32B open-weight model from sugoitoolkit for local AI inference.

Qwen3-VL-4B-Instruct-GGUF

Alibaba4B8 GB RAM

4B open-weight model from Alibaba for local AI inference.

madlad400-3b-mt

google3B4 GB RAM

3B open-weight model from google for local AI inference.

MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking-GGUF

GnLOLot1B2 GB RAM

1B open-weight model from GnLOLot for local AI inference.

svara-tts-v1

kenpathUnknown8 GB RAM

Unknown open-weight model from kenpath for local AI inference.

Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-NEO-CODE-Di-IMatrix-MAX-GGUF

DavidAU40B48 GB RAM

40B open-weight model from DavidAU for local AI inference.

Holo-3.1-35B-A3B-GGUF

Hcompany35B48 GB RAM

35B open-weight model from Hcompany for local AI inference.

gemma-4-12B-it-qat-q4_0-gguf

google12B16 GB RAM

The gemma-4-12B-it-qat-q4_0-gguf is a large language model developed by Google that features approximately 12 billion parameters. It excels at conversational interactions and handles any-to-any queries effectively, making it ideal for customer support bots or general dialogue systems. Running this model locally requires a system with sufficient RAM and

Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF

DavidAU9B16 GB RAM

9B open-weight model from DavidAU for local AI inference.

LTX-2.3-GGUF

unslothUnknown8 GB RAM

Unknown open-weight model from unsloth for local AI inference.

gemma-4-E4B-it-qat-GGUF

Google4B8 GB RAM

4B open-weight model from Google for local AI inference.

whisper-large-v3-turbo-gguf

handy-computerUnknown8 GB RAM

Unknown open-weight model from handy-computer for local AI inference.

LTX-2

DeepBeepMeepUnknown8 GB RAM

Unknown open-weight model from DeepBeepMeep for local AI inference.

Qwen3-4B-GGUF

Qwen4B8 GB RAM

4B open-weight model from Qwen for local AI inference.

Qwen3.5-0.8B-GGUF

Alibaba0.8B2 GB RAM

0.8B open-weight model from Alibaba for local AI inference.

Meta-Llama-3.1-8B-Instruct-GGUF

Meta8B16 GB RAM

8B open-weight model from Meta for local AI inference.

LFM2.5-1.2B-Instruct-GGUF

LiquidAI1.2B4 GB RAM

1.2B open-weight model from LiquidAI for local AI inference.

GLM-5.2-GGUF

unslothUnknown8 GB RAM

Unknown open-weight model from unsloth for local AI inference.

canary-180m-flash-gguf

handy-computerUnknown8 GB RAM

Unknown open-weight model from handy-computer for local AI inference.

whisper-large-v3-gguf

handy-computerUnknown8 GB RAM

Unknown open-weight model from handy-computer for local AI inference.

Qwen3-8B-GGUF

Alibaba8B16 GB RAM

8B open-weight model from Alibaba for local AI inference.

Jan-v3.5-4B-gguf

janhq4B8 GB RAM

4B open-weight model from janhq for local AI inference.

Kimi-K2.7-Code-GGUF

unslothUnknown8 GB RAM

Unknown open-weight model from unsloth for local AI inference.

Hello, Nice to meet you! 👋

Subscribe our newsletter to get the latest AI news.

Malcare WordPress Security