Skip to content
Read the original: Liquid AI Blog·Published PickAI score62/100

Liquid AI releases LFM2.5-VL-3B, a 3B vision-language model for edge devices

Original titleLFM2.5-VL-3B: A Better and Faster Vision-Language Model for the Edge

AISummary

Liquid AI released LFM2.5-VL-3B, an open-weight 3B vision-language model that it says rivals models twice its size while running faster on CPU and GPU. Benchmarks show large gains over LFM2-VL-3B, including ScreenSpot-v2 averaging 80.7, RefCOCO precision@1 rising from 57.1 to 87.9, and ToolSandbox rising from 26.4 to 59.5.

The model is available on Hugging Face and decodes 228 tokens/s on an Apple M5 Max.

AIWhy it matters

The post pairs benchmark gains with on-device and GPU throughput figures, showing how a 3B vision model trades size against speed and accuracy.

Read the original liquid.ai

Source: Liquid AI Blog · liquid.ai