SmolLM2-135M-GGUF
Unknown open-weight model from QuantFactory for local AI inference. Run SmolLM2-135M-GGUF locally with Ollama: ollama pull smollm2-135m:latest Minimum 8GB RAM.…
Unknown open-weight model from QuantFactory for local AI inference. Run SmolLM2-135M-GGUF locally with Ollama: ollama pull smollm2-135m:latest Minimum 8GB RAM.…
4B open-weight model from MaziyarPanahi for local AI inference. Run gemma-3-4b-it-GGUF locally with Ollama: ollama pull gemma-3-4b-it:4b Minimum 8GB RAM.…
7B open-weight model from MaziyarPanahi for local AI inference. Run Mistral-7B-Instruct-v0.3-GGUF locally with Ollama: ollama pull mistral-7b-instruct-v0.3:7b Minimum 8GB RAM.…
4B open-weight model from Ma7ee7 for local AI inference. Run Qwen3.8_4B_Distilled_GGUF locally with Ollama: ollama pull qwen3.8-4b-distilled-gguf:4b Minimum 8GB RAM.…
Unknown open-weight model from bartowski for local AI inference. Run XYZAILab_XYZ-Aquila-mini-GGUF locally with Ollama: ollama pull xyzailab-xyz-aquila-mini:latest Minimum 8GB RAM.…
4B open-weight model from Jackrong for local AI inference. Run Qwen3.5-4B-Claude-4.6-Opus-Reasoning-Distilled-GGUF locally with Ollama: ollama pull qwen3.5-4b-claude-4.6-opus-reasoning-distilled:4b Minimum 8GB RAM.…
Unknown open-weight model from Abiray for local AI inference. Run MiniMax-H3-Pruned-GGUF locally with Ollama: ollama pull minimax-h3-pruned:latest Minimum 8GB RAM.…
0.8B open-weight model from Alibaba for local AI inference. Run Qwen_Qwen3.5-0.8B-GGUF locally with Ollama: ollama pull qwen-qwen3.5-0.8b:0.8b Minimum 2GB RAM.…
2B open-weight model from Alibaba for local AI inference. Run Qwen_Qwen3.5-2B-GGUF locally with Ollama: ollama pull qwen-qwen3.5-2b:2b Minimum 4GB RAM.…
4B open-weight model from MaziyarPanahi for local AI inference. Run Qwen3-4B-Instruct-2507-GGUF locally with Ollama: ollama pull qwen3-4b-instruct-2507:4b Minimum 8GB RAM.…