Contents
- 0.1 Wan2.2-T2V-A14B-GGUF
- 0.2 Qwen3.8-2B-Distill-GGUF
- 0.3 Qwen3.8-4B-Distill-GGUF
- 0.4 Qwen3.8-27B-GSQ-RCO-GGUF
- 0.5 Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP-GGUF
- 0.6 glm-4-9b-chat-IMat-GGUF
- 0.7 Qwen3.8-27B-OBLITERATED
- 0.8 Ornith-1.5-397B-GGUF
- 0.9 Qwen3.8-Flash-Next-GGUF
- 0.10 Qwen3.8-27B-Heretic-Abliterated-Uncensored-GGUF
- 0.11 Qwen3.8-27B-Uncensored-HauhauCS-Aggressive-MTP-GGUF
- 0.12 Huihui-Qwen3.8-27B-abliterated-GGUF
- 0.13 Qwen3.8-27B-iMatrix-NVFP4-MTP-GGUF
- 0.14 Ornith-1.5-35B-A3B-GGUF
- 0.15 Ornith-1.5-9B-GGUF
- 0.16 Mistral-Small-24B-Instruct-2501-GGUF
- 0.17 Llama-3-8B-Instruct-32k-v0.1-GGUF
- 0.18 SmolLM2-135M-GGUF
- 0.19 Qwen3.6-35B-A3B-NVFP4-MTP-GGUF
- 0.20 Llama-3.3-70B-Instruct-GGUF
- 0.21 gemma-3-4b-it-GGUF
- 0.22 Mistral-7B-Instruct-v0.3-GGUF
- 0.23 Mixtral-8x22B-v0.1-GGUF
- 0.24 Qwen3.8_4B_Distilled_GGUF
- 1 Hello, Nice to meet you!
Wan2.2-T2V-A14B-GGUF
The Wan2.2-T2V-A14B-GGUF is a large-scale text-to-video model developed by QuantStack with a total of 14 billion parameters. It excels at generating high-quality video clips from text prompts, making it ideal for creative storytelling and content creation workflows. Running this locally requires significant GPU memory due to its size, so it is best suited for users with high-end hardware or those willing to use quantized GGUF versions for efficiency.
Qwen3.8-2B-Distill-GGUF
2B open-weight model from empero-ai for local AI inference.
Qwen3.8-4B-Distill-GGUF
Qwen3.8-4B-Distill-GGUF is a 4 billion parameter large language model provided by empero-ai designed for efficient deployment. It excels at reasoning tasks thanks to its distillation process and works well for local inference via llama.cpp. Running this quantized version locally requires minimal hardware resources, making it suitable for laptops or systems with limited GPU memory.
Qwen3.8-27B-GSQ-RCO-GGUF
27B open-weight model from ISTA-DASLab for local AI inference.
Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP-GGUF
27B open-weight model from DavidAU for local AI inference.
glm-4-9b-chat-IMat-GGUF
9B open-weight model from legraphista for local AI inference.
Qwen3.8-27B-OBLITERATED
27B open-weight model from OBLITERATUS for local AI inference.
Ornith-1.5-397B-GGUF
397B open-weight model from ornith-ai for local AI inference.
Qwen3.8-Flash-Next-GGUF
Unknown open-weight model from Alibaba for local AI inference.
Qwen3.8-27B-Heretic-Abliterated-Uncensored-GGUF
27B open-weight model from 0bserverx for local AI inference.
Qwen3.8-27B-Uncensored-HauhauCS-Aggressive-MTP-GGUF
27B open-weight model from HauhauCS for local AI inference.
Huihui-Qwen3.8-27B-abliterated-GGUF
27B open-weight model from huihui-ai for local AI inference.
Qwen3.8-27B-iMatrix-NVFP4-MTP-GGUF
27B open-weight model from cdiamond for local AI inference.
Ornith-1.5-35B-A3B-GGUF
35B open-weight model from ornith-ai for local AI inference.
Mistral-Small-24B-Instruct-2501-GGUF
24B open-weight model from MaziyarPanahi for local AI inference.
Llama-3-8B-Instruct-32k-v0.1-GGUF
8B open-weight model from MaziyarPanahi for local AI inference.
SmolLM2-135M-GGUF
Unknown open-weight model from QuantFactory for local AI inference.
Qwen3.6-35B-A3B-NVFP4-MTP-GGUF
35B open-weight model from michaelw9999 for local AI inference.
Llama-3.3-70B-Instruct-GGUF
70B open-weight model from MaziyarPanahi for local AI inference.
gemma-3-4b-it-GGUF
4B open-weight model from MaziyarPanahi for local AI inference.
Mistral-7B-Instruct-v0.3-GGUF
7B open-weight model from MaziyarPanahi for local AI inference.
Mixtral-8x22B-v0.1-GGUF
22B open-weight model from MaziyarPanahi for local AI inference.
