Google Releases Gemma 3 270M for On-Device AI
The new ultra-compact model from DeepMind is designed for efficient performance in resource-constrained environments like mobile and web.
Google DeepMind has expanded its Gemma 3 family with a new, ultra-compact model: Gemma 3 270M. This latest release focuses on bringing capable language model performance to resource-constrained environments where computational power and memory are limited.
At just 270 million parameters, the model is significantly smaller than its multi-billion parameter counterparts. This small footprint makes it a strong candidate for running directly on devices like smartphones, laptops, and web browsers, enabling applications that can operate offline and without relying on cloud-based APIs.
Despite its size, Gemma 3 270M features a modern architecture and a substantial 32,768-token context window. Google positions the model as an ideal, cost-effective base for developers and researchers looking to experiment with fine-tuning for specialized tasks.
The model is available now on Hugging Face under the custom Gemma license. This release continues Google's strategy of providing a range of open models to cater to different scales of development, from large-scale research to local, privacy-preserving applications.
Sources
- Visit
google/gemma-3-270m
Hugging Face
More in Text / LLM
Meituan Ships a Lighter, Sparser LongCat-Flash
The food-delivery giant's newest open model trims its mixture-of-experts design for more efficient inference under an MIT license.
DeepSeek Refreshes V4-Flash With New 0731 Checkpoint
The MIT-licensed mixture-of-experts model returns in an updated build shipping with FP8 weights for cheaper inference.
DeepSeek Ships V4-Flash, a 304B MoE Tuned for Agents
The latest checkpoint in DeepSeek's V4 line leans into agentic workflows while keeping the permissive MIT license.
0 comments
No comments yet. Be the first to weigh in.