Visual TL;DR. llamafile v0.10.5 released includes Updated llama.cpp core. Outdated llama.cpp solves Updated llama.cpp core. Updated llama.cpp core enables Support for big models. Support for big models like Ternary Bonsai 27B. Support for big models and Laguna-S-2.1 MoE. Support for big models leading to Efficient local AI. llamafile v0.10.5 released also features Packaging, docs upgrades.
- llamafile v0.10.5 released: new version of llamafile ships with updated core llama.cpp integration
- Outdated llama.cpp: older llamafile builds lacked support for new model architectures and quantization methods
- Updated llama.cpp core: integrates recent llama.cpp changes, enabling support for new model types
- Support for big models: now runs Ternary Bonsai 27B and Laguna-S-2.1 directly on local hardware
- Ternary Bonsai 27B: 6GB compressed Qwen3.6-27B model using ternary weights for efficiency
- Laguna-S-2.1 MoE: 118B parameter coding MoE model now accessible on consumer machines
- Efficient local AI: previously unwieldy models now fit and run efficiently on consumer-grade machines
- Packaging, docs upgrades: full details on announcement, including improved packaging and documentation
Visual TL;DR
