llama.cpp v0.6.0 adds Clef, Qwen3.8-Flash-Next, and Metal speedups
AIThe llama.cpp v0.6.0 release adds Clef support for text and vision, along with high-quality support for Qwen3.8-Flash-Next. It also brings a major Metal performance improvement and a new llama_batch_ext API, and the project website at llama.app has been refreshed.



