llama.cpp v0.6.0 adds multimodal Jev-like model support and Metal speedups
Original titlenew Llama.cpp release ships with (multimodal!) Jev-like models support, performance upgrade for Metal and more! 🔥
AISummary
llama.cpp v0.6.0 adds support for multimodal Jev-like models and a performance upgrade for Metal, according to Merve Noyan. The post says models can be served with a single command, `llama serve -hf ggml-org/Clef-Flash-GGUF`, and points to a trending list of decision models on Hugging Face.
Source: Merve Noyan · x.comPublished · added here