Qwen3.5-4B-GGUF

by Alibaba

Qwen3.5-4B-GGUF is a compact 4-billion parameter language model developed by Alibaba that supports conversational tasks and image-to-text conversion. It excels at handling lightweight inference workloads while maintaining strong performance in unsloth-optimized environments for text generation. Running this model locally requires minimal hardware resources, making it ideal for users seeking fast, efficient deployment on standard consumer devices.

Parameters 4B
RAM Required 8 GB
Context 4,096

♥ 373 people have liked this model on HuggingFace

⇩ 1,168,246 downloads on HuggingFace

How to Get This Model

Ollama ollama pull qwen3.5-4b:4b
HuggingFace View model page →
LM Studio Search "Qwen3.5-4B-GGUF" in LM Studio's Discover tab, or download the GGUF above.