by antirez
The deepseek-v4-gguf model is a quantized Mixture-of-Experts architecture provided by antirez with unknown parameter counts available in 2-bit and 4-bit formats. It excels at efficient inference for large language tasks while maintaining high performance through its specialized MoE structure, making it ideal for users needing reduced memory footprints without significant capability loss. Running this model locally requires moderate hardware resources to handle the quantized weights effectively, offering a fast and practical solution for deploying advanced AI capabilities on consumer-grade systems.
Parameters
Unknown
RAM Required
8 GB
Context
4,096
♥ 451 people have liked this model on HuggingFace
⇩ 1,525,559 downloads on HuggingFace
How to Get This Model
Ollama
ollama pull deepseek-v4:latest
HuggingFace
View model page →
LM Studio
Search "deepseek-v4-gguf" in LM Studio's Discover tab, or download the GGUF above.
