The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • Hardware estimates
  • RSS feed
  • llms.txt
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestOpenAI4
OpenAIVision-Language

H Company releases Holo4 for computer-use agents

The new vision-language model is built to drive generalist agents that operate real software interfaces.

Sep 28, 2026
NotableOther

H Company has released Holo4, a vision-language model aimed squarely at one of the hardest problems in applied AI: building agents that can actually operate a computer the way a person does. According to the company's launch post on Hugging Face, Holo4 is positioned as the engine for "generalist computer-use agents" — systems that perceive a screen and act on it.

Computer-use is a distinct discipline from general chat or coding. A model has to read a graphical interface, locate the right controls, and chain actions together reliably across many steps. That requires tight coupling between visual understanding and action planning, which is why Holo4 is framed as a vision-language model rather than a text-only system.

Why it matters

The race to make agents that can click, type, and navigate real applications has drawn in most major labs, and a credible open-weight entry changes the calculus for developers who want to build on top of this capability rather than rent it through a closed API.

  • Holo4 targets generalist use, not a single app or workflow
  • It is distributed through Hugging Face under a custom ("other") license
  • The focus is on perception-plus-action rather than conversation alone

As always, the real test will be reliability on long, multi-step tasks, where small perception errors compound quickly. For now, Holo4 adds another serious option to the growing field of models built specifically to put agents to work inside everyday software.

Sources

    OlderLiquid AI's LFM2.5-VL-DSpark targets faster VLM inferenceUnknown · Vision-Language · 2 weeks agoNewerPhonon-2 brings on-device ASR to Apple SiliconFermionResearch · Speech → Text · last week

    Get the model


    Specs

    LicenseOTHER

    Modalities

    Vision-Language
    2 versions — view changelog

    The Weekly Weights

    Every open release that mattered, one email a week.

    0 comments

    No comments yet. Be the first to weigh in.

    More from OpenAI

    All OpenAI releases →
    OpenAI/Text / LLMMulti-GPU server

    Reflection releases Beam, a 501B open-weight model

    The startup's first frontier-scale model ships with downloadable weights and a mixture-of-experts design aimed at reasoning.

    Oct 5, 2026
    OpenAI/Any-to-Any

    Thinking Machines ships Inkling, its first open model

    The Mira Murati-founded lab makes its debut with an open-weights, reasoning-focused language model.

    Jul 15, 2026
    gpt-oss-20b
    OpenAI/ReasoningRuns on a laptop

    OpenAI Releases 21B Open-Weight MoE Model

    The new `gpt-oss-20b` is an Apache 2.0-licensed Mixture-of-Experts model designed to run efficiently on consumer-grade hardware.

    Aug 4, 2025

    More in Vision-Language

    All Vision-Language →
    LFM2 d1-3B
    LiquidAI/Vision-LanguageRuns anywhere

    Liquid AI's d1-3B brings multimodal models to the edge

    The new LFM2-based d1-3B is a compact vision-language model aimed at running decisions directly on-device.

    Oct 7, 2026
    pplx-decider-v1-27b
    Perplexity Ai/Vision-LanguageConsumer GPU

    Perplexity releases a 27B model for multimodal routing

    The open-weight 'decider' model is designed to classify queries and route them inside Perplexity's stack.

    Oct 1, 2026
    Clef
    Cloudflare/Vision-LanguageConsumer GPU

    Cloudflare's Clef brings structured decisions to open models

    The new open-weight vision-language family outputs typed, structured results and arrives alongside a reinforcement-learning fine-tuning platform.

    Oct 1, 2026