gemma-4-E4B-it-GGUF

4B open-weight model from Google for local AI inference.

Run gemma-4-E4B-it-GGUF locally with Ollama: ollama pull gemma-4-e4b-it:4b

Minimum 8GB RAM. Context length: 4096 tokens.