LiquidAI releases LFM2.5-VL-3B, a 3B multimodal model for on-device use
Original titleLiquidAI/LFM2.5-VL-3B
AISummary
LiquidAI has released LFM2.5-VL-3B, a 3B-parameter multimodal model that processes text and images and is built on the LFM2.5-2.6B language model with a SigLIP2 NaFlex vision encoder.
It runs at 228 tokens/s on an Apple M5 Max and 116 tokens/s on an AMD Ryzen AI Max+ 395 in under 3.3 GB of memory, with a 32,768-token context length. The model is available in native, GGUF, ONNX and MLX formats on Hugging Face.
Source: Liquid AI · new models on Hugging Face · huggingface.coPublished · added here