gemma-4-31B-it-GGUF

31B open-weight model from Google for local AI inference.

Run gemma-4-31B-it-GGUF locally with Ollama: ollama pull gemma-4-31b-it:31b

Minimum 24GB RAM. Context length: 4096 tokens.