Liquid AI has launched LFM2.5-VL-3B, a 3.1 billion parameter vision-language model designed for on-device applications. This model excels at reading digital screens, identifying objects with coordinates, and processing documents and charts. It also supports function calling, enabling it to interact with tools based on text or image inputs. While achieving strong performance on various benchmarks, it is noted as a non-reasoning model, prioritizing low latency. AI
IMPACT Enables on-device screen reading and tool calling, potentially accelerating GUI automation and offline AI applications.
RANK_REASON New model release from a frontier AI lab. [lever_c_demoted from frontier_release: ic=2 ai=1.0]
Read on Mastodon — sigmoid.social →
- 2.6B
- LFM2.5-VL-3B
- Liquid Ai
- Apache Software License 2.0
- Apple M5 Max
- Gemma 4 E4B
- InternVL-3.5-4B
- LFM2.5-2.6B
- LFM Open License v1.0
- Qwen3.5 4B
- SigLIP2 NaFlex
AI-generated summary · Google Gemini · from 2 sources. How we write summaries →