deepseek-v4-gguf

by antirez

The deepseek-v4-gguf model is a quantized Mixture-of-Experts architecture provided by antirez with unknown parameter counts available in 2-bit and 4-bit formats. It excels at efficient inference for large language tasks while maintaining high performance through its specialized MoE structure, making it ideal for users needing reduced memory footprints without significant capability loss. Running this model locally requires moderate hardware resources to handle the quantized weights effectively, offering a fast and practical solution for deploying advanced AI capabilities on consumer-grade systems.

Parameters Unknown
RAM Required 8 GB
Context 4,096

♥ 451 people have liked this model on HuggingFace

⇩ 1,525,559 downloads on HuggingFace

How to Get This Model

Ollama ollama pull deepseek-v4:latest
HuggingFace View model page →
LM Studio Search "deepseek-v4-gguf" in LM Studio's Discover tab, or download the GGUF above.