The new v0.6.0 release packs a lot of good stuff:
- Clef support (text + vision)
- High-quality support for Qwen3.8-Flash-Next
- Massive performance improvement with Metal
- New
llama_batch_extAPI
Also the website got a nice refresh: https://llama.app
The llama.cpp v0.6.0 release adds Clef support for text and vision, along with high-quality support for Qwen3.8-Flash-Next. It also brings a major Metal performance improvement and a new llama_batch_ext API, and the project website at llama.app has been refreshed.
The new v0.6.0 release packs a lot of good stuff:
llama_batch_ext APIAlso the website got a nice refresh: https://llama.app
llama.cpp v0.6.0 https://github.com/ggml-org/llama.cpp/releases/tag/v0.6.0View quoted post on X
Source: Georgi Gerganov · x.comPublished · added here