Skip to content
View original post on X: Georgi Gerganov· 36/100AI score36/100

llama.cpp v0.6.0 adds Clef, Qwen3.8-Flash-Next, and Metal speedups

AISummary

The llama.cpp v0.6.0 release adds Clef support for text and vision, along with high-quality support for Qwen3.8-Flash-Next. It also brings a major Metal performance improvement and a new llama_batch_ext API, and the project website at llama.app has been refreshed.

Post on XView on X
@ggerganov

The new v0.6.0 release packs a lot of good stuff:

  • Clef support (text + vision)
  • High-quality support for Qwen3.8-Flash-Next
  • Massive performance improvement with Metal
  • New llama_batch_ext API

Also the website got a nice refresh: https://llama.app

ggml@ggml_org
llama.cpp v0.6.0 https://github.com/ggml-org/llama.cpp/releases/tag/v0.6.0
View quoted post on X

Source: Georgi Gerganov · x.comPublished · added here