Close Menu
NCIJ Network NCIJ Network
    What's Hot

    Triple lock move is significant, but it’s a gamble

    September 29, 2026

    Chris Mason: Andy Burnham delivers deeply political speech with a personal core

    September 29, 2026

    Away’s New Series 3 Luggage Plays It Safe—That’s the Point

    September 29, 2026
    Facebook X (Twitter) Instagram
    Trending
    • Triple lock move is significant, but it’s a gamble
    • Chris Mason: Andy Burnham delivers deeply political speech with a personal core
    • Away’s New Series 3 Luggage Plays It Safe—That’s the Point
    • OpenAI Launches dots: Always-On GPT-6 Astra Agents That Work From Their Own Cloud Computers
    • FBI tells ShinyHunters members to turn themselves in after recent arrest
    • $50 million in Bitget hacker swaps puts NEAR Intents’ ‘permissionless’ claim to the test
    • Hidden stem cells may be fueling spinal stenosis
    • Drones Are Already on the Front Lines of Wildfire Response. Robots and AI Could Be Next.
    • About
      • Our Team
      • Editorial Policy
      • Editorial Independence
      • International Support
    • Trust & Standards
      • AI Usage Policy
      • Conflict of Interest Policy
      • Corrections Policy
      • Ethics Policy
      • Fact-Checking Policy
      • Source Protection
    • Get Involved
      • Guide for Sources
      • Support Independent Journalism
    • Legal
      • Cookie Policy
      • Privacy Policy
      • Terms of Use
    Facebook X (Twitter) Instagram
    NCIJ Network NCIJ Network
    Tuesday, September 29
    • Home
    • World
    • Ai
    • Business
    • Politics
    • Health
    • Crypto
    • Science
    • Technology
    • Cybersecurity
    • Defense & Security
    • Economy
    • Energy
    • Europe
    • More
      • Fact Check
      • Investigations
      • Opinion & Analysis
      • Environment
    NCIJ Network NCIJ Network
    Home»Artificial Intelligence

    Liquid AI Releases LFM2.5-VL-3B: A 3B Vision-Language Model That Reads Screens, Grounds Objects, and Calls Tools On-Device

    NCIJ NETWNCIJ NETWORKBy NCIJ NETWNCIJ NETWORKAugust 13, 2026 Artificial Intelligence No Comments3 Mins Read
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Yesterday, Liquid AI released LFM2.5-VL-3B. It is a 3.1B-parameter vision-language model built for on-device deployment. The model reads digital screens across mobile, web, and desktop. It grounds objects to coordinates, parses documents and charts, and calls tools from text or image input. Liquid AI reports an average of 69.4 across 28 vision benchmarks. That matches InternVL-3.5-4B and sits 0.7 points behind Qwen3.5-4B, both 4.7B models. The model is non-reasoning, so it answers directly and keeps latency low. It fits in roughly 3 GB of memory and decodes 228 tokens/s on an Apple M5 Max.

    Is it deployable?

    Yes, the checkpoint ships in four formats: native, GGUF, ONNX, and MLX. Day-one runtimes include llama.cpp, MLX, vLLM, SGLang, and ONNX. It fits in roughly 3 GB of memory.

    • Which company levels: The LFM Open License v1.0 is Apache-2.0-based with one change: free commercial use ends once a company’s annual revenue reaches $10M USD. So indie developers, startups, and SMBs under that line can ship commercially at no cost. Enterprises above it must negotiate a commercial license with Liquid AI. Research, education, and non-profit use carry no revenue limit.
    • Industries: Consumer electronics, automotive, industrial and robotics, financial services, healthcare, and e-commerce. Also QA and RPA vendors that automate GUIs.
    • Applications: On-device screen agents, GUI test automation, PDF-to-structured-text with layout labels, invoice and receipt OCR, near-real-time object detection in vehicles, offline translation of menus and road signs, and multi-image comparison.

    So, What is new?

    LFM2.5-VL-3B extends LFM2-VL-3B along four axes.

    • Screen and UI understanding: The model averages 80.7 on ScreenSpot-v2 across desktop (78.7), mobile (81.2), and web (82.2). Liquid AI reports Gemma-4-E4B at 51.2 and Qwen3.5-4B at 78.5, with the larger InternVL-3.5-4B ahead at 84.1.
    • Function calling: This is new to the VL line. ToolSandbox moves from 26.4 to 59.5. BFCL v4 moves from 20.5 to 32.5. Tool calls are emitted as Pythonic calls between <|tool_call_start|> and <|tool_call_end|> tokens.
    • Grounding: RefCOCO-avg precision@1 rises from 57.1 to 87.9, a 30-point gain driven by scaled synthetic grounding data.
    • Multi-image input: BLINK improves from 50.2 to 61.5, and MuirBench from 34.9 to 58.3.

    Architecture and training

    The language backbone is LFM2.5-2.6B. The vision tower is a SigLIP2 NaFlex shape-optimized 400M encoder. NaFlex handles native resolution by splitting large images into non-overlapping 512×512 patches plus a resized whole-image thumbnail. Context length is 32,768 tokens, and 16 languages are supported.

    Pre-training used approximately 34T tokens. Vocabulary was doubled to 128K by extending the existing tokenizer in place, which improves non-Latin script coverage. Vision pre-training was scaled 4× in tokens with curated and synthetic caption, OCR, grounding, and instruction-following data.

    Post-training is SFT with knowledge distillation from a larger teacher and Antidoom training, followed by multi-reward reinforcement learning.

    The model is non-reasoning. It answers directly, which is the design choice behind its latency profile.

    Benchmarks

    Liquid AI evaluated across 28 vision benchmarks using vLLM 0.26.0 in non-reasoning mode. LFM2.5-VL-3B averages 69.4, matching InternVL-3.5-4B (69.4) and landing 0.7 points behind Qwen3.5-4B (70.1). Both comparison models are 4.7B parameters.

    Notable individual results: RealWorldQA 73.1 against InternVL-3.5-4B at 67.7, TextVQA 84.3 against Qwen3.5-4B at 81.2, MMStar 63.3, MathVista-mini 68.5, ChartQA 81.3, DocVQA 91.1, and OCRBench v1 84.2. CountBenchQA regressed to 87.3 from 92.2 in the prior release.

    On text-only evaluation, IFEval reaches 82.3, up from 72.9. Gemma-4-E4B still leads there at 87.9.

    Calls Grounds LFM2.5VL3B Liquid model Objects OnDevice Reads Releases screens Tools VisionLanguage
    NCIJ NETWNCIJ NETWORK
    • Website

    Keep Reading

    OpenAI Launches dots: Always-On GPT-6 Astra Agents That Work From Their Own Cloud Computers

    The Guardian view on Andy Burnham: grounds to hope for a better Britain | Editorial

    Nebius Opens 2026 Physical AI Awards: Five $150K Compute Credit Prizes

    Nebius Opens 2026 Physical AI Awards: Five $150K Compute Prizes, Nine Judges, and an October 25 Deadline

    Peter Brandt Says Bitcoin May Hit $600K By 2029, Calls XRP A ‘Fool Coin’

    Google Research Open-Sources RRSI: AI Agents That Improve Their Own Harness Without Overfitting

    Add A Comment
    Leave A Reply Cancel Reply

    Editors Picks

    Triple lock move is significant, but it’s a gamble

    September 29, 2026

    Chris Mason: Andy Burnham delivers deeply political speech with a personal core

    September 29, 2026

    Away’s New Series 3 Luggage Plays It Safe—That’s the Point

    September 29, 2026

    OpenAI Launches dots: Always-On GPT-6 Astra Agents That Work From Their Own Cloud Computers

    September 29, 2026
    Latest Posts

    Bitcoin collateral: MARA’s $600M Long Ridge financing

    August 7, 2026

    Truck Brake Controller’s Safety Recall Doubled as Hidden Security Fix

    August 7, 2026

    The best classic slasher movie you’ll never watch

    August 7, 2026

    Subscribe to News

    Get the latest sports news from NewsSite about world, sports and politics.

    NCIJ Network is an independent digital news platform delivering trusted investigative journalism, European and global news, in-depth analysis, and fact-based reporting with accuracy, transparency, and integrity.

    Facebook X (Twitter) Instagram Pinterest YouTube

    Triple lock move is significant, but it’s a gamble

    September 29, 2026

    Chris Mason: Andy Burnham delivers deeply political speech with a personal core

    September 29, 2026

    Away’s New Series 3 Luggage Plays It Safe—That’s the Point

    September 29, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Type above and press Enter to search. Press Esc to cancel.