19 Sources
[1]
Google's Gemma 3 is an open source, single-GPU AI with a 128K context window
Most new AI models go big -- more parameters, more tokens, more everything. Google's newest AI model has some big numbers, but it's also tuned for efficiency. Google says the Gemma 3 open source model is the best in the world for running on a single GPU or AI accelerator. The latest Gemma model is
[2]
Google calls Gemma 3 the most powerful AI model you can run on one GPU
Richard Lawler is a senior editor following news across tech, culture, policy, and entertainment. He joined The Verge in 2021 after several years covering news at Engadget. A little over a year after releasing two "open" Gemma AI models built from the same technology behind its Gemini AI, Google
[3]
Google claims Gemma 3 reaches 98% of DeepSeek's accuracy - using only one GPU
The economics of artificial intelligence have been a hot topic of late, with startup DeepSeek AI claiming eye-opening economies of scale in deploying GPU chips. Two can play that game. On Wednesday, Google announced its latest open-source large language model, Gemma 3, came close to achieving the
[4]
Google unveils Gemma 3 multi-modal AI models
Gemma 3 supports vision-language inputs and text outputs, handles context windows up to 128k tokens, and understands more than 140 languages. Google DeepMind has introduced Gemma 3, an update to the company's family of generative AI models, featuring multi-modality that allows the models to
[5]
Google announces new Gemma 3 AI models for researchers
The DOJ wants to break up Google, suggests splitting Chrome and Android Summary Gemma 3 by Google offers portability by running on a single GPU/TPU, unlike typical workstation-grade hardware. Gemma 3 shares tech with Gemini 2.0, supports 35 languages, and offers models with up to 27B parameters. It
[6]
Google unveils open source Gemma 3 model with 128k context window
Join our daily and weekly newsletters for the latest updates and exclusive content on industry-leading AI coverage. Learn More Even as large language and reasoning models remain popular, organizations increasingly turn to smaller models to run AI processes with fewer energy and cost
[7]
Google's new Gemma 3 AI models are fast, frugal, and ready for phones
Table of Contents Table of Contents Ready for mobile devices Versatile, and ready to deploy Google's AI efforts are synonymous with Gemini, which has now become an integral element of its most popular products across the Worksuite software and hardware, as well. However, the company has also
[8]
Introducing Gemma 3: The most capable model you can run on a single GPU or TPU
High performance delivered faster with quantized models: Gemma 3 introduces official quantized versions, reducing model size and computational requirements while maintaining high accuracy. For a deeper dive into the technical details behind these capabilities, as well as a comprehensive overview
[9]
Google's New AI Model Gemma 3 Shines for Creative Writers, Falls Short Elsewhere - Decrypt
On Tuesday, Google released Gemma 3, an open-source AI model based on Gemini 2.0 that packs surprising muscle for its size. The full model runs on a single GPU, yet Google benchmarks depict it as though it's competitive enough when pitted against larger models that require significantly more
[10]
Google introduces the Gemma 3 family of accessible lightweight models - SiliconANGLE
Google introduces the Gemma 3 family of accessible lightweight models Continuing a drive to make its artificial intelligence models more accessible, Google LLC today announced the next generation of its lightweight open-source family of Gemma large language models that can run on a single graphics
[11]
Google's New AI Model Outperforms DeepSeek-V3, OpenAI's o3-mini
Google on Wednesday announced Gemma 3, the next iteration in the Gemma family of open-weight models. It is a successor to the Gemma 2 model released last year. The small model comes in a range of parameter sizes - 1B, 4B, 12B and 27B. The model also supports a longer context window of 128K tokens.
[12]
Google pitches Gemma 3 as the "best single-accelerator AI"
A little over a year after the release of its initial Gemma AI models, Google has introduced Gemma 3, designed for developers creating versatile AI applications. These models support over 35 languages and can run on devices ranging from phones to workstations, with capabilities to analyze text,
[13]
Google's New Gemma 3 Open-Source AI Models Can Run on a Single GPU
Google released the Gemma 3 family of artificial intelligence (AI) models on Wednesday. Successor to the Gemma 2 series, which was introduced in August 2024, the new open-source models arrive with text and visual reasoning capabilities. The Mountain View-based tech giant said these models offer
[14]
Google's Lightweight Gemma 3 Open Model Nearly Matches DeepSeek R1
Gemma 3 models are also multimodal and support over 140 languages. Google has introduced the Gemma 3 series of open models, and they look pretty incredible, given the small size. The search giant says Gemma 3 models can be loaded on a single Nvidia H100 GPU, and it matches the performance of much
[15]
New Google Gemma 3 Multimodal AI Models Launched
Google has today released the Gemma 3 family of models, which represents a significant advancement in the field of artificial intelligence. These models introduce new improvements in multimodal capabilities, extended context handling, multilingual support, and training efficiency. Designed to
[16]
New Google Gemma 3 Multimodal AI Model Beats DeepSeek V3 : Performance Tested
Gemma 3, Google's latest suite of lightweight, open source AI models, is reshaping the landscape of artificial intelligence by emphasizing efficiency and accessibility. Despite its compact design, it delivers performance that rivals -- and often surpasses -- larger models like DeepSeek V3 and o3
[17]
Google Gemma 3 : Best Open-Weight AI Model Yet?
Google has recently launched Gemma 3, a fantastic family of open-weight, multimodal AI models designed to set new benchmarks in artificial intelligence. With model sizes ranging from 1 billion to 27 billion parameters, Gemma 3 caters to a wide array of applications, including creative writing,
[18]
Google unveils Gemma 3 lightweight AI models for all devices
Google today introduced Gemma 3, a series of advanced, lightweight open models developed using the same research behind its Gemini 2.0 models. Clement Farabet, VP of Research at Google DeepMind, described them as "our most advanced, portable, and responsibly developed open models yet." Designed to
[19]
Gemma 3: Google's New AI Beats OpenAI's o3-mini and DeepSeek-V3
Google has launched Gemma 3, the third generation of its open-source AI models. The model is better than rivals like DeepSeek-V3, o3-mini of OpenAI, and Meta's Llama 3-405B. The new Gemma 3 was launched on March 13, 2025. It builds on the success of Gemma 2 while improving efficiency, benchmark
Share
Copy Link
Google introduces Gemma 3, an open-source AI model optimized for single-GPU performance, featuring multimodal capabilities, extended context window, and improved efficiency compared to larger models.

Google has unveiled Gemma 3, the latest iteration of its open-source AI model, designed to deliver high performance on a single GPU or TPU. This release marks a significant advancement in AI efficiency and accessibility for developers and researchers
1
.Gemma 3 boasts several improvements over its predecessors:
Extended Context Window: The model now supports a 128,000-token context window, a substantial increase from the previous 8,192 tokens
1
.Multimodal Processing: Capable of handling text, high-resolution images, and short videos
2
.Language Support: Gemma 3 understands over 140 languages, with pre-trained support for 35
4
.Model Sizes: Available in 1B, 4B, 12B, and 27B parameter versions
4
.Google claims Gemma 3 achieves 98% of DeepSeek R1's accuracy while using only one GPU, compared to the estimated 32 GPUs required for R1
3
. This efficiency is attributed to:Distillation: Extracting trained weights from larger models to enhance Gemma 3's capabilities
3
.Optimization Techniques: Including quantization, improved key-value cache layouts, and GPU weight sharing
3
.Related Stories
Gemma 3 is designed for versatility across various computing environments:
On-Device Usage: Optimized for running on phones, computers, and single GPU/TPU setups
5
.Developer Tools: Available through Vertex AI, Google Colab, and downloadable via Kaggle and Hugging Face
5
.Safety Features: Includes ShieldGemma 2, an image safety checker to filter dangerous, explicit, or violent content
1
.Gemma 3's release signifies a shift towards more efficient AI models:
Democratizing AI: By reducing hardware requirements, Gemma 3 makes advanced AI more accessible to a broader range of developers and researchers
1
.Competitive Landscape: Google positions Gemma 3 as a strong competitor to models like Meta's Llama 3 and DeepSeek's R1
3
.Future of AI Efficiency: This development aligns with the industry trend towards creating more powerful yet resource-efficient AI models
3
.Summarized by
Navi
[4]
[5]
02 Apr 2026•Technology

22 May 2025•Technology

01 Aug 2024

1
Technology

2
Technology

3
Policy and Regulation
