惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

阮一峰的网络日志
阮一峰的网络日志
Google DeepMind News
Google DeepMind News
Engineering at Meta
Engineering at Meta
Threat Intelligence Blog | Flashpoint
Threat Intelligence Blog | Flashpoint
宝玉的分享
宝玉的分享
Latest news
Latest news
aimingoo的专栏
aimingoo的专栏
云风的 BLOG
云风的 BLOG
美团技术团队
V
Visual Studio Blog
F
Full Disclosure
腾讯CDC
H
Help Net Security
D
DataBreaches.Net
M
MIT News - Artificial intelligence
罗磊的独立博客
博客园 - 司徒正美
N
Netflix TechBlog - Medium
U
Unit 42
Vercel News
Vercel News
I
InfoQ
S
SegmentFault 最新的问题
B
Blog RSS Feed
博客园 - 三生石上(FineUI控件)
Microsoft Security Blog
Microsoft Security Blog
有赞技术团队
有赞技术团队
博客园_首页
The GitHub Blog
The GitHub Blog
T
Tailwind CSS Blog
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
NISL@THU
NISL@THU
P
Privacy & Cybersecurity Law Blog
Recorded Future
Recorded Future
Y
Y Combinator Blog
S
Schneier on Security
P
Proofpoint News Feed
T
Tenable Blog
Cloudbric
Cloudbric
博客园 - 【当耐特】
The Register - Security
The Register - Security
人人都是产品经理
人人都是产品经理
Google DeepMind News
Google DeepMind News
C
Cyber Attacks, Cyber Crime and Cyber Security
量子位
A
Arctic Wolf
N
News and Events Feed by Topic
H
Hacker News: Front Page
MongoDB | Blog
MongoDB | Blog
The Hacker News
The Hacker News
Webroot Blog
Webroot Blog

Developer tools

We're rolling out AlphaEvolve widely to solve Google Cloud customers' hardest problems. Expanding Managed Agents in Gemini API: background tasks, remote MCP and more The latest AI news we announced in June 2026 Ask an AI expert: What exactly is the full stack? Interactions API: our primary interface for Gemini models and agents DiffusionGemma: 4x faster text generation Bringing the latest Gemini models to Apple developers Gemma 4 QAT models: Optimizing model compression for mobile and laptop efficiency Kaggle is making AI benchmark creation effortless Introducing Gemma 4 12B: a unified, encoder-free multimodal model How we used Gemini to build Google I/O 2026 Take our I/O 2026 quiz, vibe coded in Google AI Studio. Here's what developers can do with the latest Google Play updates. Building the agentic future: Developer highlights from I/O 2026 I/O 2026 Introducing Managed Agents in the Gemini API Bring any idea to life: Google AI Studio at I/O 2026 Gemini API File Search is now multimodal: build efficient, verifiable RAG Accelerating Gemma 4: faster inference with multi-token prediction drafters The latest AI news we announced in April 2026 Reduce friction and latency for long-running jobs with Webhooks in Gemini API Join the new AI Agents Vibe Coding Course from Google and Kaggle Deep Research Max: a step change for autonomous research agents Start vibe coding in AI Studio with your Google AI subscription. Prepay for the Gemini API to get more control over your spend Introducing Learn Mode: your personal coding tutor in Google Colab Gemma 4: Byte for byte, the most capable open models New ways to balance cost and reliability in the Gemini API The latest AI news we announced in March 2026 Improve coding agents’ performance with Gemini API Docs MCP and Agent Skills. Build with Veo 3.1 Lite, our most cost-effective video generation model
See what 3 builders are making with Gemma 4
Amy Eisinger · 2026-06-10 · via Developer tools

After 150 million downloads of Gemma 4, a few creations caught our eye. Here’s how three builders are using Gemma 4 to push creative boundaries and build new apps.

"Gemma 4" is in the center with photos surrounding of a piano, lion, music, a person, and an English tutoring app interface

Your browser does not support the audio element.

Listen to article

This content is generated by Google AI. Generative AI is experimental

[[duration]] minutes

We recently released Gemma 4, our most capable open models to date. Since then, they have been downloaded more than 150 million times, and we’ve been expanding the family’s capabilities. We introduced Multi-Token Prediction (MTP) to accelerate inference, and recently released the 12B Unified model and Quantization-Aware-Training (QAT) checkpoints. Released under an Apache 2.0 license, Gemma 4 gives builders and organizations flexibility to fine-tune and deploy models across a variety of environments, from edge devices to local workstations.

Many builders are sharing what they’ve created with Gemma 4, showcasing how the models’ capabilities translate into real-world applications. Here are three highlights of what people and companies are creating.

Build low-latency, on-device apps.

The team at the app building company HubX used Gemma 4 to build BetterSpeak, an offline AI English tutoring platform. BetterSpeak uses the edge-optimized Gemma 4 E2B (effective 2B parameters) model as the reasoning engine for its on-device pipeline, enabling private, low-latency tutoring without the need for an internet connection.

To overcome mobile hardware constraints, HubX deployed the 4-bit quantized version of the model released by Google. This version handles tasks like grammar explanations and progress monitoring across multiple languages. By leveraging Gemma 4’s native audio input capabilities, the app supports direct speech-to-speech learning, reducing costs while ensuring user privacy by processing all vocal and text data entirely on-device.

A user profile screen in the English tutoring app BetterSpeak

The offline AI English tutoring platform BetterSpeak, built by HubX.

A screen in the BetterSpeak app to toggle on the Download Offline Mode

The offline AI English tutoring platform BetterSpeak, built by HubX.

A user profile screen in BetterSpeak showing the offline pack application in progress

The offline AI English tutoring platform BetterSpeak, built by HubX.

A screen showing the offline mode toggle turned on in a user's BetterSpeak profile settings

The offline AI English tutoring platform BetterSpeak, built by HubX.

A screen in the BetterSpeak app showing tutoring instruction for talking about past experiences

The offline AI English tutoring platform BetterSpeak, built by HubX.

A screen in the BetterSpeak app showing offline access to tutoring progress

The offline AI English tutoring platform BetterSpeak, built by HubX.

Get creative with vision capabilities.

Gemma 4 can perform a wide range of vision-language tasks, like object detection, visual question answering (VQA), image captioning and reasoning across multiple images.

A builder who goes by @measure_plan on X used this capability by prompting Gemma 4 to perform VQA through a specific persona. The model effectively maintained a "medieval bard" character while accurately identifying objects in the room. As the builder takes different actions, Gemma 4 stays in the persona, identifying a "glass of amber liquid" and "shelves laden with bound tomes" without breaking character.

Gamify the world around you.

Gemma 4 makes processing long-form content easy, with the larger models offering a context window of up to 256K. This expanded memory is crucial for projects like the one created by @GOROman on X, who built an app that reimagines the real world as an adventure video game. In gaming, context is everything. The large context window allows the app to remember a long history of what’s recently happened in its world.

Gemma 4 sets a new standard and offers you a chance to build locally with maximum control. Try it in Google AI Edge Gallery on iOS or Android, or explore it in Google AI Studio.

Get more stories from Google in your inbox.

Done. Just one step more.

Check your inbox to confirm your subscription.

You are already subscribed to our newsletter.

You can also subscribe with a

Related stories