惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

WordPress大学
WordPress大学
G
Google Developers Blog
小众软件
小众软件
V
V2EX
月光博客
月光博客
腾讯CDC
aimingoo的专栏
aimingoo的专栏
J
Java Code Geeks
Y
Y Combinator Blog
人人都是产品经理
人人都是产品经理
B
Blog RSS Feed
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
Microsoft Azure Blog
Microsoft Azure Blog
博客园 - 【当耐特】
D
Docker
M
MIT News - Artificial intelligence
Google DeepMind News
Google DeepMind News
N
Netflix TechBlog - Medium
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
I
InfoQ
MongoDB | Blog
MongoDB | Blog
Apple Machine Learning Research
Apple Machine Learning Research
Jina AI
Jina AI

Latest from TechRadar in News

VodafoneThree gets Ofcom approval to bring satellite connectivity to your smartphone NYT Connections today – my hints and answers for April 16 (#1040) Quordle hints and answers for Thursday, April 16 (game #1543) NYT Strands hints and answers for Thursday, April 16 (game #774) Is this the tipping point for AI at work? New Gallup survey finds half of all US employees now use it in some way Allbirds — the shoe viral company — just pivoted into AI, and I wish this were an Onion headline 'Every Apple user needs to know about this nasty scam': Fake warnings tell users their iCloud data will be… 'Makes it even more disappointing': Microsoft backs fossil fuel big time with $7 billion deal in race for AI… 'Maybe it’s not science fiction': Solar panels are causing rainwater to fall in one of the driest places… Maine becomes first US state to pass data centre construction ban Dozens of WordPress plugins hijacked to target thousands of sites Drone-killing laser weapons greenlit for use in US airspace – FAA and Defense Department say high-energy weapons are ‘ready to protect all air travelers from illicit drone use’ despite airspace restrictions and friendly-fire incidents 'We are currently being extorted' — crypto giant Kraken says it is facing extortion attack, here's… McGraw Hill becomes latest to see its Salesforce data hacked Looking for a new PC? Now might be great time to upgrade, as Gartner figures claim shipments are rising — while… Farewell Surface Hub — Microsoft kills off its super-sized touchscreen displays, but you might still be able to get one if you act fast 'We have no interest in patient data in the UK': Palantir UK head defends record as criticisms rise Amazon’s new AI Bio Discovery tool can provide ‘every researcher’ with ‘lab-in-the-loop drug discovery’ – 40+ AI biology models can filter 300,000 novel antibody candidates down to the top results for testing in just weeks Over 100 Chrome Web Store extensions found stealing user data from thousands of accounts OpenAI reveals its Mythos rival designed for cybersecurity pros NYT Connections hints and answers for Tuesday, April 14 (game #1038) Forget Dr Doolittle, study finds animals might not only want to use tech, but they also want to talk to us with it… 'The decision is deeply troubling': Tesla gets a green light for Full Self-Driving in Europe — but not… OpenAI flags third-party data issue — all macOS users should update now Microsoft says Copilot is for ‘entertainment' not work, Meta’s Muse Spark and 7 other AI stories you… Man Utd vs Leeds Live Streams: How to watch Premier League 2025/26 from anywhere in the world, team news What is the release date for Invincible season 4 episode 7 on Prime Video? Linux rules on using AI-generated code - Copilot is OK, but humans must take 'full responsibility for the… The Lenovo Legion Go 2 handheld costs more than two Nvidia RTX 5080 GPUs — and that's genuinely absurd Secretlab is launching its first Diablo desk, with a design that 'traces the infernal history' of the series
Google’s new Gemini Omni AI can turn almost anythin...
ESchwartzwri · 2026-05-20 · via Latest from TechRadar in News
Goole IO 2026 screenshot
(Image credit: Google)

  • Google introduced Gemini Omni Flash
  • It aims to make video creation easier by letting users refine projects naturally, rather than using editing software
  • It's emphasizing transparency and safety through AI watermarking and identity protections

Google’s next big AI move is aimed squarely at creativity. The company has introduced Gemini Omni at Google I/O 2026 as part of its massive slate of new Gemini features.

Omni is supposed to combine Gemini's reasoning abilities with media creation tools that can generate and edit content across different formats.

The first release, Gemini Omni Flash, focuses on video and arrives with an unusually ambitious goal. Google wants people to create content from nearly any kind of input, whether that starts with text, images, audio, or existing video.

Gemini Omni Flash is rolling out through the Gemini app, Google Flow, YouTube Shorts, and YouTube Create, with broader expansion planned later for developers and enterprise customers.

Introducing Gemini Omni: Create Anything from Anything - YouTube Introducing Gemini Omni: Create Anything from Anything - YouTube

Watch On

The announcement builds on work Google has already been doing with AI-generated visuals. In 2025, Nano Banana expanded Gemini’s image capabilities and became a surprisingly practical tool for everything from restoring aging photographs to turning rough sketches into polished concepts.

Gemini Omni is Google’s attempt to push that idea much further. The company described Gemini Omni as a way to replace tradational editing software with a conversation that can continually refine a video.

Conversational editing

One of Gemini Omni’s biggest ideas is removing complexity from editing. Google says users can modify videos through natural language while preserving consistency between changes.

Sign up for breaking news, reviews, opinion, top tech deals, and more.

Characters stay recognizable. Scenes maintain continuity. Motion remains coherent instead of resetting every time a prompt changes. The system is also designed to better understand how objects behave in the physical world, incorporating improved handling of motion, gravity, and movement dynamics.

That's how the mirror above ripples like liquid when someone touches it, or how a sculpture can be made of bubbles. Google is trying to position Gemini Omni as something larger than a video generator.

That puts Google directly into a rapidly escalating competition around AI media tools. But it's a race about who can make AI video tools feel intuitive enough that ordinary people actually want to use them, as much as anything else. Google’s answer appears to be taking the conversational route.

Eventually, Google said Gemini Omni will go beyond video. Future versions are expected to support combinations of photos, prompts, music, and reference footage into a single project.

Trusting AI creations

Powerful creative AI creates a challenge of trust, which Google acknowledged. The company is keen to highlight how videos created with Gemini Omni include SynthID watermarking technology intended to identify AI-generated media. The company also says verification tools will work across Gemini, Chrome, and Search as part of broader transparency efforts.

Users will initially be able to create video avatars based on themselves, including their own voice. But more advanced capabilities involving speech modification remain under evaluation while Google works on safety considerations.

That cautious approach reflects the increasingly awkward balancing act facing every major AI company. Building more capable systems doesn't mean trust in them will be built in tandem.


Google logo on a black background next to text reading 'Click to follow TechRadar'

Follow TechRadar on Google News and add us as a preferred source to get our expert news, reviews, and opinion in your feeds.


Purple circle with the words Best business laptops in white

Eric Hal Schwartz is a freelance writer for TechRadar with more than 15 years of experience covering the intersection of the world and technology. For the last five years, he served as head writer for Voicebot.ai and was on the leading edge of reporting on generative AI and large language models. He's since become an expert on the products of generative AI models, such as OpenAI’s ChatGPT, Anthropic’s Claude, Google Gemini, and every other synthetic media tool. His experience runs the gamut of media, including print, digital, broadcast, and live events. Now, he's continuing to tell the stories people want and need to hear about the rapidly evolving AI space and its impact on their lives. Eric is based in New York City.