Gemmable-4-12B-MTP-GGUF

12B open-weight model from Mia-AiLab for local AI inference.

Run Gemmable-4-12B-MTP-GGUF locally with Ollama: ollama pull gemmable-4-12b-mtp:12b

Minimum 16GB RAM. Context length: 4096 tokens.