llama.cpp v0.6.0 adds Clef, Qwen3.8-Flash-Next, and Metal speedups
Original titleThe new v0.6.0 release packs a lot of good stuff:
AISummary
The llama.cpp v0.6.0 release adds Clef support for text and vision, along with high-quality support for Qwen3.8-Flash-Next. It also brings a major Metal performance improvement and a new llama_batch_ext API, and the project website at llama.app has been refreshed.
Source: Georgi Gerganov · x.comPublished · added here