TII Brings Falcon to Speech With New ASR Model
The team behind the Falcon language models expands into speech-to-text with its first automatic speech recognition release.
The Technology Innovation Institute (TII), best known for its Falcon line of open language models, is extending the family into audio. In a post on Hugging Face, the group introduced Falcon ASR, its first automatic speech recognition model for turning spoken audio into text.
The release marks a notable step beyond text generation for a group that has been one of the more visible contributors to the open-weights ecosystem. Rather than a new version of an existing model, Falcon ASR is an initial entry in a new modality, pairing the Falcon brand with a speech-to-text capability that developers can build on directly.
Why it matters
Speech recognition remains one of the most practical building blocks in applied AI, underpinning transcription, voice assistants, captioning, and multimodal pipelines. A credible open option from an established lab gives teams more freedom to run and adapt ASR on their own terms.
- A first move into audio for the Falcon family
- Positioned as an openly available speech-to-text model
- Published through TII's Hugging Face presence
TII has not detailed every specification in the announcement, and prospective users should check the model card for licensing terms and intended use before deploying. Even so, Falcon ASR signals that the group intends to broaden its footprint across modalities rather than stay confined to language modeling.
Sources
- Visit
Introducing Falcon ASR
Announcement
More from NVIDIA
All NVIDIA releases →Falcon-Emirati tunes an LLM for local dialect
TII's Falcon family gets a variant built around Emirati Arabic, aiming at culture and nuance rather than generic Gulf Arabic.
NVIDIA's Nemotron-3 Brings Streaming Speaker Diarization
The new open model applies NVIDIA's Sortformer approach to identify who's speaking, in real time.
NVIDIA's Kumo Takes Aim at Tabular Prediction
A new foundation model targets structured data, where spreadsheets and databases still dominate real-world machine learning.
More in Speech → Text
All Speech → Text →Cactus Compute's Whistle brings speech-to-text to the edge
A compact on-device ASR model built for edge hardware and WebAssembly, with early support for English, German, and French.
Phonon-2 brings on-device ASR to Apple Silicon
A low-bit quantized, Parakeet-based speech recognizer built to run locally on Mac hardware.
Audio8-ASR-Infinite brings streaming bilingual speech recognition
A new open model targets real-time transcription for Chinese and English audio.
0 comments
No comments yet. Be the first to weigh in.