by prism-ml
Bonsai-27B-gguf is a compact 27-billion parameter language model developed by prism-ml that utilizes quantization for efficient deployment. It excels at conversational tasks and general reasoning while running smoothly on llama.cpp with support for both CPU and CUDA hardware acceleration. Users can run this model locally on modest hardware, though performance will vary depending on whether they utilize 1-bit quantization or have access to a GPU.
Parameters
27B
RAM Required
24 GB
Context
4,096
♥ 783 people have liked this model on HuggingFace
⇩ 1,432,039 downloads on HuggingFace
How to Get This Model
Ollama
ollama pull bonsai-27b:27b
HuggingFace
View model page →
LM Studio
Search "Bonsai-27B-gguf" in LM Studio's Discover tab, or download the GGUF above.
