惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Google DeepMind News
Google DeepMind News
博客园 - 聂微东
Vercel News
Vercel News
aimingoo的专栏
aimingoo的专栏
F
Fortinet All Blogs
Microsoft Security Blog
Microsoft Security Blog
MongoDB | Blog
MongoDB | Blog
B
Blog
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
WordPress大学
WordPress大学
Apple Machine Learning Research
Apple Machine Learning Research
阮一峰的网络日志
阮一峰的网络日志
大猫的无限游戏
大猫的无限游戏
GbyAI
GbyAI
Martin Fowler
Martin Fowler
M
MIT News - Artificial intelligence
The GitHub Blog
The GitHub Blog
博客园_首页
博客园 - 叶小钗
腾讯CDC
G
Google Developers Blog
Blog — PlanetScale
Blog — PlanetScale
宝玉的分享
宝玉的分享
D
Docker

PetaPixel

The Camp Snap 2 Has More Photo Accessories and is Smaller Than Ever Polaroid Go Generation 3 Is the World’s Smallest Instant Analog Camera China Lucky’s New Color C200 Film Has Arrived in the US and Looks Great Chinese Private Equity Firm HSG Emerges as Leading Bidder to Buy Stake in Leica Camera Leica Cine Compact 1 Projector Promises a Plug-and-Play Home Theater Experience There Are No Photos Allowed at the Studio Ghibli Amusement Park Xiaomi 17T Pro Review: High-End Photography Without the Flagship Price Photographer Builds the iPhone Camera App He Always Wanted New $50 Y2K-Inspired Compact Camera Looks Thin In Form and Function Save Big on Macro Photography Essentials Nvidia’s New Chip Aims to Upend the Creative Laptop Market Japanese Zoo Considering Photo Ban After US Tourists Invade Punch the Monkey’s Enclosure Years in the Making, Glass Imaging Is Delivering on its Promise to Transform Smartphone Photography Gen Z are Five Times More Likely to Have a Plan for Photos After Death Long Island Beach Where Marilyn Monroe Posed for Iconic Photos Honored With Plaque Camera Trap Captures One-Armed Gorilla Raising Newborn in the Wild Photographer’s Camera Gear Gets Jet Washed by 20-Foot Wave in Tahiti Photojournalists Say ICE Agents Targeted Them and Their Cameras at Delaney Hall Protests Model Sues Fashion Brand After it AI-Generated Pictures of Her Easy Macro Photography Tips for Incredible Close-Up Photos The Myth of Intent in Photography Samsung’s Competitors Have a Better Samsung Camera Than Samsung Does Thypoch Simera 50mm f/1.4 Review: Crisp and Clean The AI Film ‘Dreams of Violets’ Is How You Get Me to Hate Movies Photographer Documents the Vanishing Wildlife of the ‘American Amazon’ Press Photographer Releases 30-Year Archive of Iconic Celebrity Images Photographer Granted Rare Access to Cambridge’s May Balls for 40 Years Atmos Is a Weather App By Photographers, for Photographers A Feature-Length AI-Generated ‘Live Action’ Movie Is Premiering at Tribeca for Some Reason Despite Being a Member, YoloLiv Isn’t Complying with the Micro Four Thirds Standard
Google’s New Gemini Omni AI Video Model Can Do Crazy Things
Jeremy Gray · 2026-05-28 · via PetaPixel

Google’s new Gemini Omni artificial intelligence (AI) model can do some wild things. The model’s key promise is to create anything from, well, anything.

Google says its new Gemini Omni model can “create anything from any input,” including audio, video, photos, and text. The model starts with video generation, which users can then edit via conversational text with Gemini. This first model, Gemini Omni Flash, is launching now in the Gemini app, Google Flow, and YouTube Shorts.

As Google explains, editing AI-generated video using text is straightforward. The model also promises to keep things consistent after editing, including characters, and Omni can remember what was visible in previous scenes.

Prompt: Make the sculpture out of bubbles.

The company even promises that Gemini Omni can use its “intuitive understanding of physics,” effectively “bridging the gap from photorealism to meaningful storytelling.”

Prompt: A marble rolling fast on a chain reaction style track, continuous smooth shot.

Users have already achieved impressive results with Gemini Omni. For example, ex-Google product manager Bilawal Sidhu gave Gemini Omni a photo with a sketched drone path on it and had the AI generate drone POV footage.

Gave google omni a sketched camera path and asked it to generate drone POV footage. pic.twitter.com/cQZFMtOkEi

— Bilawal Sidhu (@bilawalsidhu) May 26, 2026

The Verge‘s Allison Johnson calls Omni “wild,” and had the AI bring her child’s stuffed animal, Buddy, to life. Buddy went on exciting AI adventures, including white-water rafting and snowboarding.

“The results are such a mixed bag they’re baffling. Some were very good — much more consistent and true to my prompt than when I was testing out Veo five months ago,” Johnson writes. “But even the best clips Omni cooked up for me still have certain AI jump scares, like when Buddy suddenly switches orientation while he’s skydiving.”

Prompt: turn this into realistic footage, using the drawing only as a guide for movement, do not show the drawing in the final video

As Johnson tested, Omni’s biggest claim to fame, being able to combine a wide variety of input media with AI-generated video, veers from technologically impressive to potentially hazardous. One of her deepfakes even convinced her husband, “a man who has looked at me in real life basically every single day for the last decade.”

Whether this is neat or terrifying depends on who is asked.

“I can’t be the only one to think, that this just has no reason to exist,” writes near_photography on Threads in response to Johnson’s post above. “There is no net benefit to society from this capability.”

Prompt: Apply the pose and motion from input video to provided character from this image. Apply style from image reference to the new video

As Google notes, all videos generated using Omni include its “imperceptible SynthID digital watermark,” which makes it easy for users to confirm if something was made with Google’s AI inside Gemini, Gemini in Chrome, and Google Search. But what if someone isn’t using those platforms?

Google is bringing this technology directly into YouTube Shorts and YouTube Create, for example, and it’s impossible to predict what people will do with it there.


Image credits: Google