Bonsai 2 27B
Bonsai 2 27B Articles
Browse 3 articles about Bonsai 2 27B.

Bonsai 2 27B: A 27B Model That Runs in Under 6GB on a Laptop
Bonsai 2 27B uses ternary quantization to shrink a 27B model to under 6GB while keeping 98% of FP16 performance. Here's how it works.
Bonsai 2 27Bternary quantizationMLX 2-bit model

Run Bonsai 2 27B Locally on a Mac: Ternary Quantization Explained
Bonsai 2 27B compresses a 27B reasoning model to 8.6GB with ternary weights, hitting ~47 tok/s on an M5 Max MacBook via MLX.
Bonsai 2 27Bternary quantization LLMrun 27B model on Mac

How to Run Bonsai 2 27B Locally: Full Install Guide
Install and serve Bonsai 2 27B with llama.cpp locally. Covers VRAM usage, slot configuration, GGUF formats, and real-world test results.
Bonsai 2 27Brun Bonsai locallyllama.cpp ternary model