The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestFermionResearch2
FermionResearchSpeech → Text

Phonon-2 brings on-device ASR to Apple Silicon

A low-bit quantized, Parakeet-based speech recognizer built to run locally on Mac hardware.

Sep 28, 2026
UpdateOther
Phonon-2

FermionResearch has published Phonon-2, a compact automatic speech recognition model designed to run locally on Apple Silicon. The model is built on NVIDIA's Parakeet ASR architecture and shipped in a low-bit quantized form, with roughly 600 million parameters and a focus on English transcription.

The pitch here is practical rather than flashy: instead of streaming audio to a cloud endpoint, Phonon-2 aims to do the work on the device itself. For developers building Mac-native apps — note-takers, captioning tools, voice interfaces — that means lower latency, no per-request API cost, and audio that never leaves the machine.

Why it matters

Parakeet has earned a reputation as one of the stronger open ASR families, but running it well on consumer hardware has usually meant trimming it down. Phonon-2's low-bit quantization is the lever that makes a sub-gigabyte footprint feasible on Apple's unified-memory chips.

  • Architecture derived from NVIDIA's Parakeet ASR line
  • About 0.6B parameters, quantized to low precision
  • English-only, optimized for on-device Apple Silicon inference

The release is distributed under a non-standard license, so teams planning commercial use should read the terms on the model card before shipping. For anyone who wants fast, private transcription without a server bill, it's a notable addition to the on-device toolkit.

Sources

  • FermionResearch/Phonon-2

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Parameters600M
LanguagesEnglish
Architectureparakeet_tdt_five_value
LicenseOTHER
Downloads2.1K
Likes152

Modalities

Speech → Text

0 comments

No comments yet. Be the first to weigh in.

More in Speech → Text

Audio8-ASR-Infinite
Edge0/Speech → Text

Audio8-ASR-Infinite brings streaming bilingual speech recognition

A new open model targets real-time transcription for Chinese and English audio.

Sep 21, 2026
Parakeet Redux
moondream/Speech → Text

Moondream shrinks Parakeet ASR for CPUs

A ternary-quantized take on the Parakeet TDT speech model aims to run transcription without a GPU.

Sep 18, 2026
Unknown/Text → Speech

Nari Labs Ships Qwen3-Based TTS and ASR Models

The startup pairs speech synthesis and recognition built on Qwen3, pitching accuracy, low latency, and lower cost.

Sep 14, 2026