Loading market data...
ai

Liquid AI Releases LFM2.5-VL-3B: A 3B Vision-Language Model That Reads Screens, Grounds Objects, and Calls Tools On-Device

MarkTechPost
Read Full Article at MarkTechPost
Share:PostShare
Ad Slot — In-Article (728x90)

Liquid AI released LFM2. 5-VL-3B, a 3. 1B-parameter vision-language model built for on-device deployment. It averages 80. 7 on ScreenSpot-v2 and lifts RefCOCO grounding from 57. 1 to 87. 9. Function calling is new to the VL line, with ToolSandbox moving from 26. 4 to 59. 5.

The model fits in roughly 3 GB and decodes 228 tokens/s on an Apple M5 Max. The post Liquid AI Releases LFM2. 5-VL-3B: A 3B Vision-Language Model That Reads Screens, Grounds Objects, and Calls Tools On-Device appeared first on MarkTechPost.

This is a summary. For the full story, read the original article at MarkTechPost.

Original source: MarkTechPost

Ad Slot — Below Article (300x250)