by Alibaba
Qwen3.5-9B-GGUF is a compact large language model developed by Alibaba containing 9 billion parameters. It excels at conversational tasks and image-to-text conversion, making it ideal for lightweight applications that require efficient region US deployment. Running this model locally is practical on consumer-grade hardware thanks to its small footprint, offering fast inference speeds even on modest GPUs when paired with optimization libraries like Unsloth.
Parameters
9B
RAM Required
16 GB
Context
4,096
♥ 840 people have liked this model on HuggingFace
⇩ 1,229,858 downloads on HuggingFace
How to Get This Model
Ollama
ollama pull qwen3.5-9b:9b
HuggingFace
View model page →
LM Studio
Search "Qwen3.5-9B-GGUF" in LM Studio's Discover tab, or download the GGUF above.
