Qwen3.5-4B-GGUF

4B open-weight model from Alibaba for local AI inference.

Run Qwen3.5-4B-GGUF locally with Ollama: ollama pull qwen3.5-4b:4b

Minimaal 8 GB RAM. Contextlengte: 4096 tokens.