New to LM Studio? Read the full setup guide.
Contents
- 1 AI Models for LM Studio
- 1.0.1 Wan2.2-T2V-A14B-GGUF
- 1.0.2 Qwen3.8-2B-Distill-GGUF
- 1.0.3 Qwen3.8-4B-Distill-GGUF
- 1.0.4 Qwen3.8-27B-GSQ-RCO-GGUF
- 1.0.5 Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP-GGUF
- 1.0.6 glm-4-9b-chat-IMat-GGUF
- 1.0.7 Qwen3.8-27B-OBLITERATED
- 1.0.8 Ornith-1.5-397B-GGUF
- 1.0.9 Qwen3.8-Flash-Next-GGUF
- 1.0.10 Qwen3.8-27B-Heretic-Abliterated-Uncensored-GGUF
- 1.0.11 Qwen3.8-27B-Uncensored-HauhauCS-Aggressive-MTP-GGUF
- 1.0.12 Huihui-Qwen3.8-27B-abliterated-GGUF
- 1.0.13 Qwen3.8-27B-iMatrix-NVFP4-MTP-GGUF
- 1.0.14 Ornith-1.5-35B-A3B-GGUF
- 1.0.15 Ornith-1.5-9B-GGUF
- 1.0.16 Mistral-Small-24B-Instruct-2501-GGUF
- 1.0.17 Llama-3-8B-Instruct-32k-v0.1-GGUF
- 1.0.18 SmolLM2-135M-GGUF
- 1.0.19 Qwen3.6-35B-A3B-NVFP4-MTP-GGUF
- 1.0.20 Llama-3.3-70B-Instruct-GGUF
- 1.0.21 gemma-3-4b-it-GGUF
- 1.0.22 Mistral-7B-Instruct-v0.3-GGUF
- 1.0.23 Mixtral-8x22B-v0.1-GGUF
- 1.0.24 Qwen3.8_4B_Distilled_GGUF
- 1.0.25 Meta-Llama-3-8B-Instruct-GGUF
- 1.0.26 Melody1437-26B-A4B-v2.0-GGUF
- 1.0.27 Serenity-26B-A4B-GGUF
- 1.0.28 Dark-Scarlett-v0.3-26B-A4B-GGUF
- 1.0.29 XYZAILab_XYZ-Aquila-mini-GGUF
- 1.0.30 Qwen3.5-4B-Claude-4.6-Opus-Reasoning-Distilled-GGUF
- 1.0.31 Parable-Granite-4.1-8B-Claude-Fable-5-GGUF
- 1.0.32 Qwen3.6-14B-A3B-FableVibes-GGUF
- 1.0.33 MiniMax-H3-Pruned-GGUF
- 1.0.34 Qwen_Qwen3.5-0.8B-GGUF
- 1.0.35 Qwen3.6-35B-A3B-APEX-GGUF
- 1.0.36 supergemma4-26b-uncensored-gguf-v2
- 1.0.37 gemma-4-26B-A4B-it-APEX-GGUF
- 1.0.38 Qwopus3.6-27B-Coder-MTP-GGUF
- 1.0.39 Qwen_Qwen3.5-2B-GGUF
- 1.0.40 Qwen3.5-397B-A17B-GGUF
- 1.0.41 Qwen3-4B-Instruct-2507-GGUF
- 1.0.42 Ternary-Bonsai-8B-gguf
- 1.0.43 PinkCherry_NSFW_LTX23
- 1.0.44 Parable-Granite-4.1-3B-Claude-Fable-5-GGUF
- 1.0.45 Qwen_Qwen3.5-35B-A3B-GGUF
- 1.0.46 endless-frontier_BigBang-v1-GGUF
- 1.0.47 Qwen_Qwen3.5-4B-GGUF
- 1.0.48 Qwen3.6-27B-Heretic-Uncensored-FINETUNE-NEO-CODE-Di-IMatrix-MAX-GGUF
- 1.0.49 Qwen3.5-9B-DeepSeek-V4-Flash-GGUF
- 1.0.50 Parable-Qwen3-8B-Claude-Fable-5-GGUF
- 1.0.51 grug-27b-gguf
- 1.0.52 GLM-4.7-Flash-GGUF
- 1.0.53 gemma-4-31B-it-qat-q4_0-gguf
- 1.0.54 Qwen3.5-27B-GGUF
- 1.0.55 Kwaipilot_KAT-Coder-V2.5-Dev-GGUF
- 1.0.56 Parable-Qwen3-4B-Claude-Fable-5-GGUF
- 1.0.57 gemma-4-E2B-it-qat-GGUF
- 1.0.58 ACE-Step-1.5-GGUF
- 1.0.59 Huihui-DeepSeek-V4-Flash-0731-abliterated-GGUF
- 1.0.60 POCKET-35B-GGUF
- 1.0.61 gemma-4-26B-A4B-it-qat-q4_0-gguf
- 1.0.62 gemma-4-E2B-it-qat-q4_0-gguf
- 1.0.63 gemma-4-E4B-it-qat-q4_0-gguf
- 1.0.64 automotive
- 1.0.65 DeepSeek-V4-Flash-0731-GGUF
- 1.0.66 LFM2.5-2.6B-GGUF
- 1.0.67 Qwen3.5-122B-A10B-GGUF
- 1.0.68 Qwen3.8-27B-Uncensored-GGUF
- 1.0.69 MiniMax-H3
- 1.0.70 MiniMax-H3_GGUFs
- 1.0.71 audio.cpp-gguf
- 1.0.72 Inkling-Small-GGUF
- 1.0.73 maple-preview-GGUF
- 1.0.74 KAT-Coder-V2.5-Dev-APEX-GGUF
- 1.0.75 Qwen3-TTS-GGUF
- 1.0.76 MiniMax-H3-GGUF
- 1.0.77 Muse-Glimmer-30B-GGUF
- 1.0.78 Laguna-XS-2.1-APEX-GGUF
- 1.0.79 Qwen3-30B-A3B-Thinking-2507-GGUF
- 1.0.80 Qwen3.8-27B-GGUF
- 1.0.81 Higgs-Audio-v3-Studio
- 1.0.82 DeepSeek-V4-Flash-GGUF
- 1.0.83 Qwen3.5-2B-GGUF
- 1.0.84 Qwen3.5-4B-Uncensored-HauhauCS-Aggressive
- 1.0.85 Qwen3.5-35B-A3B-GGUF
- 1.0.86 Llama-3.2-3B-Instruct-GGUF
- 1.0.87 gemma-3-1b-it-GGUF
- 1.0.88 Qwen3.5-122B-A10B-MTP-GGUF
- 1.0.89 FLUX.2-klein-4B-GGUF
- 1.0.90 Qwen2.5-7B-Instruct-GGUF
- 1.0.91 Qwen3-30B-A3B-GGUF
- 1.0.92 Qwen3-32B-GGUF
- 1.0.93 Qwen3-1.7B-GGUF
- 1.0.94 Kimi-K3-GGUF
- 1.0.95 Qwen2.5-0.5B-Instruct-GGUF
- 1.0.96 Qwen3-14B-GGUF
- 1.0.97 Qwen3-0.6B-GGUF
- 1.0.98 Laguna-S-2.1-GGUF
- 1.0.99 Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V7-GGUF
- 1.0.100 gemma-gguf
- 1.0.101 Qwen2.5-VL-7B-Instruct-GGUF
- 1.0.102 Qwen3-235B-A22B-GGUF
- 1.0.103 Voxtral-Small-24B-2507-gguf
- 1.0.104 Sugoi-14B-Ultra-GGUF
- 1.0.105 Qwen3-ASR-1.7B-gguf
- 1.0.106 Qwen3-30B-A3B-Instruct-2507-GGUF
- 1.0.107 Qwen2.5-3B-Instruct-GGUF
- 1.0.108 gemma-4-12b-heretic-abliterated-GGUF
- 1.0.109 Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive
- 1.0.110 Qwen2.5-Coder-7B-Instruct-GGUF
- 1.0.111 Qwen-Image-Edit-2511-GGUF
- 1.0.112 Qwen3-Coder-Next-GGUF
- 1.0.113 Qwen2.5-1.5B-Instruct-GGUF
- 1.0.114 Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-GGUF
- 1.0.115 paw-programs
- 1.0.116 Qwen2.5-Coder-32B-Instruct-GGUF
- 1.0.117 gemma-4-31B-it-qat-GGUF
- 1.0.118 gemma-4-26B-A4B-it-ultra-uncensored-heretic-i1-GGUF
- 1.0.119 gemma-4-12B-coder-fable5-composer2.5-v1-GGUF
- 1.0.120 inkling-GGUF
- 1.0.121 granite-4.1-3b-GGUF
- 1.0.122 Sugoi-32B-Ultra-GGUF
- 1.0.123 Qwen3-VL-4B-Instruct-GGUF
- 1.0.124 madlad400-3b-mt
- 1.0.125 MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking-GGUF
- 1.0.126 svara-tts-v1
- 1.0.127 Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-NEO-CODE-Di-IMatrix-MAX-GGUF
- 1.0.128 Holo-3.1-35B-A3B-GGUF
- 1.0.129 gemma-4-12B-it-qat-q4_0-gguf
- 1.0.130 Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF
- 1.0.131 LTX-2.3-GGUF
- 1.0.132 gemma-4-E4B-it-qat-GGUF
- 1.0.133 whisper-large-v3-turbo-gguf
- 1.0.134 LTX-2
- 1.0.135 Qwen3-4B-GGUF
- 1.0.136 Qwen3.5-0.8B-GGUF
- 1.0.137 Meta-Llama-3.1-8B-Instruct-GGUF
- 1.0.138 LFM2.5-1.2B-Instruct-GGUF
- 1.0.139 GLM-5.2-GGUF
- 1.0.140 canary-180m-flash-gguf
- 1.0.141 whisper-large-v3-gguf
- 1.0.142 Qwen3-8B-GGUF
- 1.0.143 Jan-v3.5-4B-gguf
- 1.0.144 Kimi-K2.7-Code-GGUF
- 1.0.145 MiniCPM5-1B-Claude-Opus-Fable5-Thinking-GGUF
- 1.0.146 Llama-3.2-1B-Instruct-Q8_0-GGUF
- 1.0.147 Bielik-11B-v3.0-Instruct-awq
- 1.0.148 Wan2.2-I2V-A14B-GGUF
- 1.0.149 Gemma-4-31B-it-abliterated
- 1.0.150 Qwen3.6-27B-Uncensored-HauhauCS-Aggressive
- 1.0.151 ThinkingCap-Qwen3.6-27B-GGUF
- 1.0.152 Wan2.1
- 1.0.153 Voxtral-Mini-4B-Realtime-2602-gguf
- 1.0.154 Qwen_Qwen3.6-35B-A3B-GGUF
- 1.0.155 Qwythos-9B-v2-GGUF
- 1.0.156 gemma-4-12B-agentic-fable5-composer2.5-v2-3.5x-tau2-GGUF
- 1.0.157 Sulphur-2-base
- 1.0.158 whisper-medium-gguf
- 1.0.159 gemma-4-E2B-it-GGUF
- 1.0.160 parakeet-tdt-0.6b-v3-gguf
- 1.0.161 Qwen3.5-9B-Uncensored-HauhauCS-Aggressive
- 1.0.162 gemma-4-26B-A4B-it-qat-GGUF
- 1.0.163 gpt-oss-20b-GGUF
- 1.0.164 Qwen3-VL-8B-Instruct-abliterated-GGUF
- 1.0.165 Flux2-Klein-9B-True-V2
- 1.0.166 Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF
- 1.0.167 Qwen3-Coder-30B-A3B-Instruct-GGUF
- 1.0.168 gemma-4-31B-it-GGUF
- 1.0.169 Gemma-4-E4B-Uncensored-HauhauCS-Aggressive
- 1.0.170 Qwopus3.6-27B-Coder-Compat-MTP-GGUF
- 1.0.171 UI-TARS-1.5-7B-GGUF
- 1.0.172 Huihui-DeepSeek-V4-Flash-abliterated-ds4-GGUF
- 1.0.173 gemma-4-12B-it-qat-GGUF
- 1.0.174 Ternary-Bonsai-27B-gguf
- 1.0.175 Gemmable-4-12B-MTP-GGUF
- 1.0.176 Qwopus3.6-35B-A3B-Coder-MTP-GGUF
- 1.0.177 Qwen3.6-35B-A3B-MTP-GGUF
- 1.0.178 gemma-4-12b-it-GGUF
- 1.0.179 Qwen3-VL-30B-A3B-Instruct-GGUF
- 1.0.180 models-moved
- 1.0.181 cohere-transcribe-03-2026-gguf
- 1.0.182 Qwen-AgentWorld-35B-A3B-GGUF
- 1.0.183 vntl-llama3-8b-v2-gguf
- 1.0.184 gemma-4-E4B-it-GGUF
- 1.0.185 Qwen3.6-27B-GGUF
- 1.0.186 Qwen3.6-35B-A3B-GGUF
- 1.0.187 Hy3-GGUF
- 1.0.188 HyperCLOVAX-SEED-Text-Instruct-1.5B-Q4_K_M-GGUF
- 1.0.189 Qwen3.5-9B-GGUF
- 1.0.190 Qwen3.5-4B-GGUF
- 1.0.191 parakeet-unified-en-0.6b-gguf
- 1.0.192 nemotron-3.5-asr-streaming-0.6b-gguf
- 1.0.193 gemma-4-26B-A4B-it-GGUF
- 1.0.194 Bonsai-27B-gguf
- 1.0.195 Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive
- 1.0.196 Qwythos-9B-Claude-Mythos-5-1M-GGUF
- 1.0.197 Ornith-1.0-35B-GGUF
- 1.0.198 Qwen3.6-27B-MTP-GGUF
- 1.0.199 deepseek-v4-gguf
- 1.0.200 Ornith-1.0-9B-GGUF
- 1.1 Hello, Nice to meet you!
AI Models for LM Studio
Every model in this directory is GGUF format, so it loads directly in LM Studio. Here is the full catalog with direct HuggingFace links.
Wan2.2-T2V-A14B-GGUF
The Wan2.2-T2V-A14B-GGUF is a large-scale text-to-video model developed by QuantStack with a total of 14 billion parameters. It excels at generating high-quality video clips from text prompts, making it ideal for creative storytelling and content creation workflows. Running this locally requires significant GPU memory due to its size, so it is best suited for users with high-end hardware or those willing to use quantized GGUF versions for efficiency.
Download for LM Studio →Qwen3.8-2B-Distill-GGUF
2B open-weight model from empero-ai for local AI inference.
Download for LM Studio →Qwen3.8-4B-Distill-GGUF
Qwen3.8-4B-Distill-GGUF is a 4 billion parameter large language model provided by empero-ai designed for efficient deployment. It excels at reasoning tasks thanks to its distillation process and works well for local inference via llama.cpp. Running this quantized version locally requires minimal hardware resources, making it suitable for laptops or systems with limited GPU memory.
Download for LM Studio →Qwen3.8-27B-GSQ-RCO-GGUF
27B open-weight model from ISTA-DASLab for local AI inference.
Download for LM Studio →Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP-GGUF
27B open-weight model from DavidAU for local AI inference.
Download for LM Studio →glm-4-9b-chat-IMat-GGUF
9B open-weight model from legraphista for local AI inference.
Download for LM Studio →Qwen3.8-27B-OBLITERATED
27B open-weight model from OBLITERATUS for local AI inference.
Download for LM Studio →Ornith-1.5-397B-GGUF
397B open-weight model from ornith-ai for local AI inference.
Download for LM Studio →Qwen3.8-Flash-Next-GGUF
Unknown open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwen3.8-27B-Heretic-Abliterated-Uncensored-GGUF
27B open-weight model from 0bserverx for local AI inference.
Download for LM Studio →Qwen3.8-27B-Uncensored-HauhauCS-Aggressive-MTP-GGUF
27B open-weight model from HauhauCS for local AI inference.
Download for LM Studio →Huihui-Qwen3.8-27B-abliterated-GGUF
27B open-weight model from huihui-ai for local AI inference.
Download for LM Studio →Qwen3.8-27B-iMatrix-NVFP4-MTP-GGUF
27B open-weight model from cdiamond for local AI inference.
Download for LM Studio →Ornith-1.5-35B-A3B-GGUF
35B open-weight model from ornith-ai for local AI inference.
Download for LM Studio →Ornith-1.5-9B-GGUF
9B open-weight model from ornith-ai for local AI inference.
Download for LM Studio →Mistral-Small-24B-Instruct-2501-GGUF
24B open-weight model from MaziyarPanahi for local AI inference.
Download for LM Studio →Llama-3-8B-Instruct-32k-v0.1-GGUF
8B open-weight model from MaziyarPanahi for local AI inference.
Download for LM Studio →SmolLM2-135M-GGUF
Unknown open-weight model from QuantFactory for local AI inference.
Download for LM Studio →Qwen3.6-35B-A3B-NVFP4-MTP-GGUF
35B open-weight model from michaelw9999 for local AI inference.
Download for LM Studio →Llama-3.3-70B-Instruct-GGUF
70B open-weight model from MaziyarPanahi for local AI inference.
Download for LM Studio →gemma-3-4b-it-GGUF
4B open-weight model from MaziyarPanahi for local AI inference.
Download for LM Studio →Mistral-7B-Instruct-v0.3-GGUF
7B open-weight model from MaziyarPanahi for local AI inference.
Download for LM Studio →Mixtral-8x22B-v0.1-GGUF
22B open-weight model from MaziyarPanahi for local AI inference.
Download for LM Studio →Qwen3.8_4B_Distilled_GGUF
4B open-weight model from Ma7ee7 for local AI inference.
Download for LM Studio →Meta-Llama-3-8B-Instruct-GGUF
8B open-weight model from MaziyarPanahi for local AI inference.
Download for LM Studio →Melody1437-26B-A4B-v2.0-GGUF
26B open-weight model from ReadyArt for local AI inference.
Download for LM Studio →Serenity-26B-A4B-GGUF
26B open-weight model from ReadyArt for local AI inference.
Download for LM Studio →Dark-Scarlett-v0.3-26B-A4B-GGUF
26B open-weight model from ReadyArt for local AI inference.
Download for LM Studio →XYZAILab_XYZ-Aquila-mini-GGUF
Unknown open-weight model from bartowski for local AI inference.
Download for LM Studio →Qwen3.5-4B-Claude-4.6-Opus-Reasoning-Distilled-GGUF
4B open-weight model from Jackrong for local AI inference.
Download for LM Studio →Parable-Granite-4.1-8B-Claude-Fable-5-GGUF
8B open-weight model from AnkitAI for local AI inference.
Download for LM Studio →Qwen3.6-14B-A3B-FableVibes-GGUF
14B open-weight model from tvall43 for local AI inference.
Download for LM Studio →MiniMax-H3-Pruned-GGUF
Unknown open-weight model from Abiray for local AI inference.
Download for LM Studio →Qwen_Qwen3.5-0.8B-GGUF
0.8B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwen3.6-35B-A3B-APEX-GGUF
35B open-weight model from mudler for local AI inference.
Download for LM Studio →supergemma4-26b-uncensored-gguf-v2
26B open-weight model from Jiunsong for local AI inference.
Download for LM Studio →gemma-4-26B-A4B-it-APEX-GGUF
26B open-weight model from mudler for local AI inference.
Download for LM Studio →Qwopus3.6-27B-Coder-MTP-GGUF
27B open-weight model from Jackrong for local AI inference.
Download for LM Studio →Qwen_Qwen3.5-2B-GGUF
2B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwen3.5-397B-A17B-GGUF
397B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwen3-4B-Instruct-2507-GGUF
4B open-weight model from MaziyarPanahi for local AI inference.
Download for LM Studio →Ternary-Bonsai-8B-gguf
8B open-weight model from prism-ml for local AI inference.
Download for LM Studio →PinkCherry_NSFW_LTX23
Unknown open-weight model from SexGod1979 for local AI inference.
Download for LM Studio →Parable-Granite-4.1-3B-Claude-Fable-5-GGUF
3B open-weight model from AnkitAI for local AI inference.
Download for LM Studio →Qwen_Qwen3.5-35B-A3B-GGUF
35B open-weight model from Alibaba for local AI inference.
Download for LM Studio →endless-frontier_BigBang-v1-GGUF
Unknown open-weight model from bartowski for local AI inference.
Download for LM Studio →Qwen_Qwen3.5-4B-GGUF
4B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwen3.6-27B-Heretic-Uncensored-FINETUNE-NEO-CODE-Di-IMatrix-MAX-GGUF
27B open-weight model from DavidAU for local AI inference.
Download for LM Studio →Qwen3.5-9B-DeepSeek-V4-Flash-GGUF
9B open-weight model from Jackrong for local AI inference.
Download for LM Studio →Parable-Qwen3-8B-Claude-Fable-5-GGUF
8B open-weight model from AnkitAI for local AI inference.
Download for LM Studio →grug-27b-gguf
27B open-weight model from ProCreations for local AI inference.
Download for LM Studio →GLM-4.7-Flash-GGUF
Unknown open-weight model from unsloth for local AI inference.
Download for LM Studio →gemma-4-31B-it-qat-q4_0-gguf
31B open-weight model from google for local AI inference.
Download for LM Studio →Qwen3.5-27B-GGUF
27B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Kwaipilot_KAT-Coder-V2.5-Dev-GGUF
Unknown open-weight model from bartowski for local AI inference.
Download for LM Studio →Parable-Qwen3-4B-Claude-Fable-5-GGUF
4B open-weight model from AnkitAI for local AI inference.
Download for LM Studio →gemma-4-E2B-it-qat-GGUF
2B open-weight model from Google for local AI inference.
Download for LM Studio →ACE-Step-1.5-GGUF
Unknown open-weight model from Serveurperso for local AI inference.
Download for LM Studio →Huihui-DeepSeek-V4-Flash-0731-abliterated-GGUF
Unknown open-weight model from huihui-ai for local AI inference.
Download for LM Studio →POCKET-35B-GGUF
35B open-weight model from FINAL-Bench for local AI inference.
Download for LM Studio →gemma-4-26B-A4B-it-qat-q4_0-gguf
26B open-weight model from google for local AI inference.
Download for LM Studio →gemma-4-E2B-it-qat-q4_0-gguf
2B open-weight model from google for local AI inference.
Download for LM Studio →gemma-4-E4B-it-qat-q4_0-gguf
4B open-weight model from google for local AI inference.
Download for LM Studio →automotive
Unknown open-weight model from flywheel-ai for local AI inference.
Download for LM Studio →DeepSeek-V4-Flash-0731-GGUF
Unknown open-weight model from DeepSeek for local AI inference.
Download for LM Studio →LFM2.5-2.6B-GGUF
2.6B open-weight model from LiquidAI for local AI inference.
Download for LM Studio →Qwen3.5-122B-A10B-GGUF
122B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwen3.8-27B-Uncensored-GGUF
27B open-weight model from JonathanColetti for local AI inference.
Download for LM Studio →MiniMax-H3
Unknown open-weight model from DeepBeepMeep for local AI inference.
Download for LM Studio →MiniMax-H3_GGUFs
Unknown open-weight model from realrebelai for local AI inference.
Download for LM Studio →audio.cpp-gguf
Unknown open-weight model from audio-cpp for local AI inference.
Download for LM Studio →Inkling-Small-GGUF
Unknown open-weight model from unsloth for local AI inference.
Download for LM Studio →maple-preview-GGUF
Unknown open-weight model from deepgrove for local AI inference.
Download for LM Studio →KAT-Coder-V2.5-Dev-APEX-GGUF
Unknown open-weight model from mudler for local AI inference.
Download for LM Studio →Qwen3-TTS-GGUF
Unknown open-weight model from Serveurperso for local AI inference.
Download for LM Studio →MiniMax-H3-GGUF
Unknown open-weight model from Abiray for local AI inference.
Download for LM Studio →Muse-Glimmer-30B-GGUF
30B open-weight model from unsloth for local AI inference.
Download for LM Studio →Laguna-XS-2.1-APEX-GGUF
Unknown open-weight model from mudler for local AI inference.
Download for LM Studio →Qwen3-30B-A3B-Thinking-2507-GGUF
30B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwen3.8-27B-GGUF
27B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Higgs-Audio-v3-Studio
Unknown open-weight model from drbaph for local AI inference.
Download for LM Studio →DeepSeek-V4-Flash-GGUF
Unknown open-weight model from DeepSeek for local AI inference.
Download for LM Studio →Qwen3.5-2B-GGUF
2B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwen3.5-4B-Uncensored-HauhauCS-Aggressive
4B open-weight model from HauhauCS for local AI inference.
Download for LM Studio →Qwen3.5-35B-A3B-GGUF
35B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Llama-3.2-3B-Instruct-GGUF
3B open-weight model from Meta for local AI inference.
Download for LM Studio →gemma-3-1b-it-GGUF
1B open-weight model from ggml-org for local AI inference.
Download for LM Studio →Qwen3.5-122B-A10B-MTP-GGUF
122B open-weight model from Alibaba for local AI inference.
Download for LM Studio →FLUX.2-klein-4B-GGUF
4B open-weight model from unsloth for local AI inference.
Download for LM Studio →Qwen2.5-7B-Instruct-GGUF
7B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwen3-30B-A3B-GGUF
30B open-weight model from MaziyarPanahi for local AI inference.
Download for LM Studio →Qwen3-32B-GGUF
32B open-weight model from MaziyarPanahi for local AI inference.
Download for LM Studio →Qwen3-1.7B-GGUF
1.7B open-weight model from MaziyarPanahi for local AI inference.
Download for LM Studio →Kimi-K3-GGUF
Unknown open-weight model from unsloth for local AI inference.
Download for LM Studio →Qwen2.5-0.5B-Instruct-GGUF
0.5B open-weight model from Qwen for local AI inference.
Download for LM Studio →Qwen3-14B-GGUF
14B open-weight model from MaziyarPanahi for local AI inference.
Download for LM Studio →Qwen3-0.6B-GGUF
0.6B open-weight model from MaziyarPanahi for local AI inference.
Download for LM Studio →Laguna-S-2.1-GGUF
Unknown open-weight model from unsloth for local AI inference.
Download for LM Studio →Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V7-GGUF
35B open-weight model from LuffyTheFox for local AI inference.
Download for LM Studio →gemma-gguf
Unknown open-weight model from rahul7star for local AI inference.
Download for LM Studio →Qwen2.5-VL-7B-Instruct-GGUF
7B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwen3-235B-A22B-GGUF
235B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Voxtral-Small-24B-2507-gguf
24B open-weight model from handy-computer for local AI inference.
Download for LM Studio →Sugoi-14B-Ultra-GGUF
14B open-weight model from sugoitoolkit for local AI inference.
Download for LM Studio →Qwen3-ASR-1.7B-gguf
1.7B open-weight model from handy-computer for local AI inference.
Download for LM Studio →Qwen3-30B-A3B-Instruct-2507-GGUF
30B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwen2.5-3B-Instruct-GGUF
3B open-weight model from Qwen for local AI inference.
Download for LM Studio →gemma-4-12b-heretic-abliterated-GGUF
12B open-weight model from culturerevolt for local AI inference.
Download for LM Studio →Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive
122B open-weight model from HauhauCS for local AI inference.
Download for LM Studio →Qwen2.5-Coder-7B-Instruct-GGUF
7B open-weight model from Qwen for local AI inference.
Download for LM Studio →Qwen-Image-Edit-2511-GGUF
Unknown open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwen3-Coder-Next-GGUF
Unknown open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwen2.5-1.5B-Instruct-GGUF
1.5B open-weight model from Qwen for local AI inference.
Download for LM Studio →Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-GGUF
35B open-weight model from LuffyTheFox for local AI inference.
Download for LM Studio →paw-programs
Unknown open-weight model from programasweights for local AI inference.
Download for LM Studio →Qwen2.5-Coder-32B-Instruct-GGUF
32B open-weight model from Qwen for local AI inference.
Download for LM Studio →gemma-4-31B-it-qat-GGUF
31B open-weight model from Google for local AI inference.
Download for LM Studio →gemma-4-26B-A4B-it-ultra-uncensored-heretic-i1-GGUF
26B open-weight model from Google for local AI inference.
Download for LM Studio →gemma-4-12B-coder-fable5-composer2.5-v1-GGUF
12B open-weight model from yuxinlu1 for local AI inference.
Download for LM Studio →inkling-GGUF
Unknown open-weight model from unsloth for local AI inference.
Download for LM Studio →granite-4.1-3b-GGUF
3B open-weight model from ibm-granite for local AI inference.
Download for LM Studio →Sugoi-32B-Ultra-GGUF
32B open-weight model from sugoitoolkit for local AI inference.
Download for LM Studio →Qwen3-VL-4B-Instruct-GGUF
4B open-weight model from Alibaba for local AI inference.
Download for LM Studio →madlad400-3b-mt
3B open-weight model from google for local AI inference.
Download for LM Studio →MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking-GGUF
1B open-weight model from GnLOLot for local AI inference.
Download for LM Studio →svara-tts-v1
Unknown open-weight model from kenpath for local AI inference.
Download for LM Studio →Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-NEO-CODE-Di-IMatrix-MAX-GGUF
40B open-weight model from DavidAU for local AI inference.
Download for LM Studio →Holo-3.1-35B-A3B-GGUF
35B open-weight model from Hcompany for local AI inference.
Download for LM Studio →gemma-4-12B-it-qat-q4_0-gguf
The gemma-4-12B-it-qat-q4_0-gguf is a large language model developed by Google that features approximately 12 billion parameters. It excels at conversational interactions and handles any-to-any queries effectively, making it ideal for customer support bots or general dialogue systems. Running this model locally requires a system with sufficient RAM and
Download for LM Studio →Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF
9B open-weight model from DavidAU for local AI inference.
Download for LM Studio →LTX-2.3-GGUF
Unknown open-weight model from unsloth for local AI inference.
Download for LM Studio →gemma-4-E4B-it-qat-GGUF
4B open-weight model from Google for local AI inference.
Download for LM Studio →whisper-large-v3-turbo-gguf
Unknown open-weight model from handy-computer for local AI inference.
Download for LM Studio →LTX-2
Unknown open-weight model from DeepBeepMeep for local AI inference.
Download for LM Studio →Qwen3-4B-GGUF
4B open-weight model from Qwen for local AI inference.
Download for LM Studio →Qwen3.5-0.8B-GGUF
0.8B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Meta-Llama-3.1-8B-Instruct-GGUF
8B open-weight model from Meta for local AI inference.
Download for LM Studio →LFM2.5-1.2B-Instruct-GGUF
1.2B open-weight model from LiquidAI for local AI inference.
Download for LM Studio →GLM-5.2-GGUF
Unknown open-weight model from unsloth for local AI inference.
Download for LM Studio →canary-180m-flash-gguf
Unknown open-weight model from handy-computer for local AI inference.
Download for LM Studio →whisper-large-v3-gguf
Unknown open-weight model from handy-computer for local AI inference.
Download for LM Studio →Qwen3-8B-GGUF
8B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Jan-v3.5-4B-gguf
4B open-weight model from janhq for local AI inference.
Download for LM Studio →Kimi-K2.7-Code-GGUF
Unknown open-weight model from unsloth for local AI inference.
Download for LM Studio →MiniCPM5-1B-Claude-Opus-Fable5-Thinking-GGUF
1B open-weight model from GnLOLot for local AI inference.
Download for LM Studio →Llama-3.2-1B-Instruct-Q8_0-GGUF
1B open-weight model from hugging-quants for local AI inference.
Download for LM Studio →Bielik-11B-v3.0-Instruct-awq
11B open-weight model from speakleash for local AI inference.
Download for LM Studio →Wan2.2-I2V-A14B-GGUF
14B open-weight model from QuantStack for local AI inference.
Download for LM Studio →Gemma-4-31B-it-abliterated
31B open-weight model from paperscarecrow for local AI inference.
Download for LM Studio →Qwen3.6-27B-Uncensored-HauhauCS-Aggressive
27B open-weight model from HauhauCS for local AI inference.
Download for LM Studio →ThinkingCap-Qwen3.6-27B-GGUF
27B open-weight model from bottlecapai for local AI inference.
Download for LM Studio →Wan2.1
Unknown open-weight model from DeepBeepMeep for local AI inference.
Download for LM Studio →Voxtral-Mini-4B-Realtime-2602-gguf
4B open-weight model from handy-computer for local AI inference.
Download for LM Studio →Qwen_Qwen3.6-35B-A3B-GGUF
35B open-weight model from Alibaba for local AI inference.
Download for LM Studio →Qwythos-9B-v2-GGUF
9B open-weight model from empero-ai for local AI inference.
Download for LM Studio →gemma-4-12B-agentic-fable5-composer2.5-v2-3.5x-tau2-GGUF
12B open-weight model from yuxinlu1 for local AI inference.
Download for LM Studio →Sulphur-2-base
Unknown open-weight model from SulphurAI for local AI inference.
Download for LM Studio →whisper-medium-gguf
Unknown open-weight model from handy-computer for local AI inference.
Download for LM Studio →gemma-4-E2B-it-GGUF
2B open-weight model from Google for local AI inference.
Download for LM Studio →parakeet-tdt-0.6b-v3-gguf
0.6B open-weight model from handy-computer for local AI inference.
Download for LM Studio →Qwen3.5-9B-Uncensored-HauhauCS-Aggressive
9B open-weight model from HauhauCS for local AI inference.
Download for LM Studio →gemma-4-26B-A4B-it-qat-GGUF
26B open-weight model from Google for local AI inference.
Download for LM Studio →gpt-oss-20b-GGUF
20B open-weight model from unsloth for local AI inference.
Download for LM Studio →Qwen3-VL-8B-Instruct-abliterated-GGUF
Qwen3-VL-8B-Instruct-abliterated-GGUF is an 8-billion parameter vision-language model developed by Alibaba that has been fully abliterated to remove proprietary constraints while retaining core capabilities. This model excels at handling complex visual reasoning and multi-step tasks within a conversational framework, making it ideal for open-ended analysis and creative generation in the US region. Running this GGUF quantized version locally is highly practical for users with mid-range GPUs, offering fast inference speeds without requiring specialized enterprise hardware.
Download for LM Studio →Flux2-Klein-9B-True-V2
9B open-weight model from wikeeyang for local AI inference.
Download for LM Studio →Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF
The Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF is a 27-billion parameter large language model created by DavidAU that combines multiple fine-tuning techniques including Unsloth and Heretic methods. This model excels at generating uncensored, highly creative, and unrestricted content across diverse topics, making it ideal for users seeking an abliterated assistant capable of multi-stage tuned responses without safety filters. Running this GGUF quantized version locally requires a substantial GPU with ample VRAM to handle the 27B parameter load efficiently, offering best performance on systems equipped with high-end hardware for fast inference speeds.
Download for LM Studio →Qwen3-Coder-30B-A3B-Instruct-GGUF
Qwen3-Coder-30B-A3B-Instruct-GGUF is a 30-billion parameter coding model developed by Alibaba that integrates advanced instruction tuning for specialized tasks. It excels at complex code generation, debugging, and conversational programming assistance, making it ideal for developers seeking high-performance solutions in the US region. Running this model locally requires substantial GPU memory to handle its large parameter count, so it is best suited for users with powerful hardware who prioritize raw coding capability over speed.
Download for LM Studio →gemma-4-31B-it-GGUF
31B open-weight model from Google for local AI inference.
Download for LM Studio →Gemma-4-E4B-Uncensored-HauhauCS-Aggressive
Gemma-4-E4B-Uncensored-HauhauCS-Aggressive is a 4-billion parameter multimodal model developed by HauhauCS that integrates advanced capabilities from Gemma 4 with ablated safety filters. This aggressive variant excels at unrestricted text generation, vision analysis, and audio processing, making it ideal for creative tasks requiring raw output without content moderation. Running this model locally demands significant GPU memory to handle its multimodal inputs, and users should expect high computational costs when processing complex audio or video streams.
Download for LM Studio →Qwopus3.6-27B-Coder-Compat-MTP-GGUF
27B open-weight model from Jackrong for local AI inference.
Download for LM Studio →UI-TARS-1.5-7B-GGUF
7B open-weight model from mradermacher for local AI inference.
Download for LM Studio →Huihui-DeepSeek-V4-Flash-abliterated-ds4-GGUF
Unknown open-weight model from huihui-ai for local AI inference.
Download for LM Studio →gemma-4-12B-it-qat-GGUF
12B open-weight model from Google for local AI inference.
Download for LM Studio →Ternary-Bonsai-27B-gguf
27B open-weight model from prism-ml for local AI inference.
Download for LM Studio →Gemmable-4-12B-MTP-GGUF
12B open-weight model from Mia-AiLab for local AI inference.
Download for LM Studio →Qwopus3.6-35B-A3B-Coder-MTP-GGUF
35B open-weight model from Jackrong for local AI inference.
Download for LM Studio →Qwen3.6-35B-A3B-MTP-GGUF
35B open-weight model from Alibaba for local AI inference.
Download for LM Studio →gemma-4-12b-it-GGUF
12B open-weight model from Google for local AI inference.
Download for LM Studio →Qwen3-VL-30B-A3B-Instruct-GGUF
30B open-weight model from Qwen for local AI inference.
Download for LM Studio →models-moved
The models-moved collection is provided by ggml-org and currently lists an unknown parameter count for its various model files. These models excel at general-purpose tasks within the US region and are best suited for users seeking accessible inference options from this specific provider. Running them locally requires downloading the appropriate GGUF quantization files, with performance varying significantly based on your GPU or CPU hardware capabilities.
Download for LM Studio →cohere-transcribe-03-2026-gguf
Unknown open-weight model from handy-computer for local AI inference.
Download for LM Studio →Qwen-AgentWorld-35B-A3B-GGUF
35B open-weight model from Alibaba for local AI inference.
Download for LM Studio →vntl-llama3-8b-v2-gguf
The vntl-llama3-8b-v2-gguf is an 8B parameter language model developed by lmg-anon that specializes in high-quality translation and conversational tasks. It excels at processing the VNTL-v5-1k dataset to deliver fluent responses, making it ideal for multilingual chat applications and localized content generation. Running this model locally requires a GPU with sufficient VRAM to handle its 8B parameter weight efficiently, ensuring responsive performance for real-time dialogue.
Download for LM Studio →gemma-4-E4B-it-GGUF
The gemma-4-E4B-it-GGUF is a 4-billion parameter language model developed by Google and optimized for Italian using GGUF quantization. It excels at handling complex reasoning tasks within the Gemma 4 architecture while supporting efficient image-to-text conversions through specialized unsloth optimizations. Running this model locally requires a modern GPU to leverage its speed, making it ideal for users seeking high-performance local inference without cloud dependencies.
Download for LM Studio →Qwen3.6-27B-GGUF
Qwen3.6-27B-GGUF is a 27-billion parameter large language model developed by Alibaba that supports advanced conversational AI and image-to-text capabilities. It excels at complex reasoning tasks, multilingual communication, and visual analysis, making it ideal for professional applications requiring high accuracy in text generation and understanding. Running this model locally typically demands substantial GPU memory and a powerful processor to handle its 27B parameters efficiently, so it is best suited for users with dedicated hardware or access to cloud instances.
Download for LM Studio →Qwen3.6-35B-A3B-GGUF
Qwen3.6-35B-A3B-GGUF is a 35-billion parameter large language model developed by Alibaba that features specialized optimizations for image-to-text conversion and multilingual conversation. This model excels at handling complex visual reasoning tasks and maintaining coherent dialogue, making it ideal for applications requiring deep contextual understanding across diverse topics. Running this model locally requires substantial GPU memory to handle its full parameter count, though quantized GGUF versions can offer a practical balance between speed and performance on high-end consumer hardware.
Download for LM Studio →Hy3-GGUF
Unknown open-weight model from vcruz305 for local AI inference.
Download for LM Studio →HyperCLOVAX-SEED-Text-Instruct-1.5B-Q4_K_M-GGUF
HyperCLOVAX-SEED-Text-Instruct-1.5B-Q4_K_M-GGUF is a 1.5 billion parameter language model created by rippertnt and optimized for use with llama-cpp. It excels at conversational tasks and general text generation, making it ideal for lightweight chat applications that require efficient inference on standard hardware. Running this model locally is practical even on modest systems due to its quantized format, offering fast response times without the need for high-end GPUs.
Download for LM Studio →Qwen3.5-9B-GGUF
Qwen3.5-9B-GGUF is a compact large language model developed by Alibaba containing 9 billion parameters. It excels at conversational tasks and image-to-text conversion, making it ideal for lightweight applications that require efficient region US deployment. Running this model locally is practical on consumer-grade hardware thanks to its small footprint, offering fast inference speeds even on modest GPUs when paired with optimization libraries like Unsloth.
Download for LM Studio →Qwen3.5-4B-GGUF
Qwen3.5-4B-GGUF is a compact 4-billion parameter language model developed by Alibaba that supports conversational tasks and image-to-text conversion. It excels at handling lightweight inference workloads while maintaining strong performance in unsloth-optimized environments for text generation. Running this model locally requires minimal hardware resources, making it ideal for users seeking fast, efficient deployment on standard consumer devices.
Download for LM Studio →parakeet-unified-en-0.6b-gguf
0.6B open-weight model from handy-computer for local AI inference.
Download for LM Studio →nemotron-3.5-asr-streaming-0.6b-gguf
0.6B open-weight model from handy-computer for local AI inference.
Download for LM Studio →gemma-4-26B-A4B-it-GGUF
The gemma-4-26B-A4B-it-GGUF is a 26-billion parameter large language model developed by Google that leverages advanced instruction tuning for high-performance reasoning. This model excels at complex text generation and multimodal tasks, making it ideal for applications requiring deep contextual understanding and precise instruction following. Running this model locally demands substantial GPU memory and significant compute power, so it is best suited for users with high-end hardware or those utilizing optimized inference frameworks like Unsloth to manage its resource requirements.
Download for LM Studio →Bonsai-27B-gguf
Bonsai-27B-gguf is a compact 27-billion parameter language model developed by prism-ml that utilizes quantization for efficient deployment. It excels at conversational tasks and general reasoning while running smoothly on llama.cpp with support for both CPU and CUDA hardware acceleration. Users can run this model locally on modest hardware, though performance will vary depending on whether they utilize 1-bit quantization or have access to a GPU.
Download for LM Studio →Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive
Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive is a 35-billion parameter large language model developed by HauhauCS that combines MoE architecture with advanced multimodal and vision capabilities. It excels at complex reasoning, multilingual image-text analysis, and generating uncensored content for aggressive or unrestricted use cases where standard safety filters are undesirable. Running this model locally requires substantial high-end GPU memory to handle its sparse mixture-of-experts structure, making it best suited for powerful workstations or servers with significant VRAM available.
Download for LM Studio →Qwythos-9B-Claude-Mythos-5-1M-GGUF
9B open-weight model from empero-ai for local AI inference.
Download for LM Studio →Ornith-1.0-35B-GGUF
Ornith-1.0-35B-GGUF is a 35-billion parameter language model developed by ornith-ai specifically for conversational tasks within the US region. It excels at maintaining natural dialogue flows and handling context-aware interactions, making it ideal for chatbots and virtual assistants. Running this model locally requires substantial GPU memory to handle its size efficiently, so it is best suited for users with high-end hardware or those willing to use quantized GGUF versions to reduce resource demands.
Download for LM Studio →Qwen3.6-27B-MTP-GGUF
Qwen3.6-27B-MTP-GGUF is a 27-billion parameter large language model developed by Alibaba that supports advanced conversational tasks and image-to-text conversion. It excels at handling complex reasoning and multi-turn dialogues, making it ideal for applications requiring deep contextual understanding and visual analysis. Running this model locally typically demands high-end GPU hardware to manage its substantial memory footprint, though quantized GGUF versions can offer a practical balance between speed and performance on consumer-grade systems.
Download for LM Studio →deepseek-v4-gguf
The deepseek-v4-gguf model is a quantized Mixture-of-Experts architecture provided by antirez with unknown parameter counts available in 2-bit and 4-bit formats. It excels at efficient inference for large language tasks while maintaining high performance through its specialized MoE structure, making it ideal for users needing reduced memory footprints without significant capability loss. Running this model locally requires moderate hardware resources to handle the quantized weights effectively, offering a fast and practical solution for deploying advanced AI capabilities on consumer-grade systems.
Download for LM Studio →Ornith-1.0-9B-GGUF
Ornith-1.0-9B-GGUF is a 9-billion parameter language model developed by ornith-ai designed for conversational interactions within the US region. It excels at generating natural dialogue and handling casual chat tasks, making it ideal for lightweight virtual assistants or simple customer support bots. Running this model locally requires modest hardware resources, allowing it to operate quickly on consumer-grade GPUs without needing massive data centers.
Download for LM Studio →
