惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Application and Cybersecurity Blog
Application and Cybersecurity Blog
The Register - Security
The Register - Security
V
Visual Studio Blog
aimingoo的专栏
aimingoo的专栏
Stack Overflow Blog
Stack Overflow Blog
IT之家
IT之家
量子位
C
Check Point Blog
博客园 - 【当耐特】
小众软件
小众软件
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
雷峰网
雷峰网
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
Microsoft Azure Blog
Microsoft Azure Blog
CTFtime.org: upcoming CTF events
CTFtime.org: upcoming CTF events
Engineering at Meta
Engineering at Meta
Recorded Future
Recorded Future
The Last Watchdog
The Last Watchdog
博客园 - Franky
N
Netflix TechBlog - Medium
Webroot Blog
Webroot Blog
A
About on SuperTechFans
cs.AI updates on arXiv.org
cs.AI updates on arXiv.org
W
WeLiveSecurity
D
Docker
S
Security Affairs
T
The Blog of Author Tim Ferriss
F
Fortinet All Blogs
Blog — PlanetScale
Blog — PlanetScale
V2EX - 技术
V2EX - 技术
Jina AI
Jina AI
Help Net Security
Help Net Security
L
LangChain Blog
P
Proofpoint News Feed
The Cloudflare Blog
WordPress大学
WordPress大学
Google DeepMind News
Google DeepMind News
Schneier on Security
Schneier on Security
Recent Announcements
Recent Announcements
Attack and Defense Labs
Attack and Defense Labs
云风的 BLOG
云风的 BLOG
V
Vulnerabilities – Threatpost
Microsoft Security Blog
Microsoft Security Blog
H
Heimdal Security Blog
P
Proofpoint News Feed
O
OpenAI News
H
Help Net Security
cs.CV updates on arXiv.org
cs.CV updates on arXiv.org
爱范儿
爱范儿
Security Archives - TechRepublic
Security Archives - TechRepublic

Replicate's blog

How to make remarkable videos with Seedance 2.0 – Replicate blog How to prompt Seedream 5.0 – Replicate blog Recraft V4: image generation with design taste – Replicate blog Run Isaac 0.1 on Replicate – Replicate blog Run FLUX.2 on Replicate – Replicate blog How to prompt Nano Banana Pro – Replicate blog Retro Diffusion's pixel art models are now on Replicate – Replicate blog Replicate is joining Cloudflare – Replicate blog Extract text from documents and images with Datalab Marker and OCR – Replicate blog How to prompt Veo 3.1 – Replicate blog IBM's Granite 4.0 is now on Replicate – Replicate blog Which image editing model should I use? – Replicate blog Introducing our new search API – Replicate blog Torch compile caching for inference speed – Replicate blog Announcing Replicate's remote MCP server – Replicate blog How to prompt Veo 3 with images – Replicate blog Open source video is back – Replicate blog Generate consistent characters – Replicate blog Bria is now on Replicate – Replicate blog How we optimized FLUX.1 Kontext [dev] – Replicate blog Compare AI video models – Replicate blog The FLUX.1 Kontext hackathon – Replicate blog How to prompt Veo 3 for the best results – Replicate blog Get the most from Google Veo 3 – Replicate blog FLUX.1 Kontext from the community – Replicate blog Use FLUX.1 Kontext to edit images with words – Replicate blog Generate incredible images with Google's Imagen 4 – Replicate blog Run OpenAI’s latest models on Replicate – Replicate blog NVIDIA H100 GPUs are here – Replicate blog Run 30,000+ LoRAs on Hugging Face with Replicate – Replicate blog Ideogram 3.0 on Replicate – Replicate blog Run MiniMax Speech-02 models with an API – Replicate blog Easel AI is now on Replicate – Replicate blog Stylized video with Wan2.1 – Replicate blog Creative roundup: avatars, lightsabers, and LoRA tricks – Replicate blog Wan2.1: generate videos with an API – Replicate blog Wan2.1 parameter sweep – Replicate blog You can now fine-tune open-source video models – Replicate blog Generate short videos with the Replicate playground – Replicate blog AI video is having its Stable Diffusion moment – Replicate blog FLUX fine-tunes are now fast – Replicate blog FLUX.1 Tools – Control and steerability for FLUX – Replicate blog NVIDIA L40S GPUs are here – Replicate blog Ideogram v2 is an outstanding new inpainting model – Replicate blog Stable Diffusion 3.5 is here – Replicate blog FLUX is fast and it's open source – Replicate blog FLUX1.1 [pro] is here – Replicate blog Using synthetic training data to improve Flux finetunes – Replicate blog Fine-tune FLUX.1 with an API – Replicate blog Fine-tune FLUX.1 to create images of yourself – Replicate blog Replicate Intelligence #12 – Replicate blog Replicate Intelligence #11 – Replicate blog Fine-tune FLUX.1 with your own images – Replicate blog Replicate Intelligence #10 – Replicate blog FLUX.1: First Impressions – Replicate blog Replicate Intelligence #9 – Replicate blog Run FLUX with an API – Replicate blog Replicate Intelligence #8 – Replicate blog Run Meta Llama 3.1 405B with an API – Replicate blog Replicate Intelligence #7 – Replicate blog Replicate Intelligence #6 – Replicate blog Replicate Intelligence #5 – Replicate blog How to get the best results from Stable Diffusion 3 – Replicate blog Run Stable Diffusion 3 on your Apple Silicon Mac – Replicate blog Push a custom version of Stable Diffusion 3 – Replicate blog Replicate Intelligence #4 – Replicate blog Run Stable Diffusion 3 on your own machine with ComfyUI – Replicate blog H100s are coming to Replicate – Replicate blog Run Stable Diffusion 3 with an API – Replicate blog Replicate Intelligence #3 – Replicate blog Replicate Intelligence #2 – Replicate blog Replicate Intelligence #1 – Replicate blog Shared network vulnerability disclosure – Replicate blog Run Snowflake Arctic with an API – Replicate blog Run Meta Llama 3 with an API – Replicate blog Run Code Llama 70B with an API – Replicate blog How to create an AI narrator for your life – Replicate blog Clone your voice using open-source models – Replicate blog Businesses are building on open-source AI – Replicate blog How to run Yi chat models with an API – Replicate blog Scaffold Replicate apps with one command – Replicate blog Using open-source models for faster and cheaper text embeddings – Replicate blog Generate music from chord progressions and text prompts with MusicGen-Chord – Replicate blog Generate images in one second on your Mac using a latent consistency model – Replicate blog How to use retrieval augmented generation with ChromaDB and Mistral – Replicate blog Fine-tune MusicGen to generate music in any style – Replicate blog Jet-setting with Llama 2 + Grammars – Replicate blog How to run Mistral 7B with an API – Replicate blog Make smooth AI generated videos with AnimateDiff and an interpolator – Replicate blog Fine-tuned models now boot in less than one second – Replicate blog Painting with words: a history of text-to-image AI – Replicate blog We're cutting our prices in half – Replicate blog A guide to prompting Llama 2 – Replicate blog Streaming output for language models – Replicate blog Fine-tune SDXL with your own images – Replicate blog Run Llama 2 with an API – Replicate blog Run SDXL with an API – Replicate blog A comprehensive guide to running Llama 2 locally – Replicate blog Fine-tune Llama 2 on Replicate – Replicate blog What happened with Llama 2 in the last 24 hours? 🦙 – Replicate blog
How to prompt Grok Imagine Video 1.5 – Replicate blog
2026-05-21 · via Replicate's blog

Grok Imagine Video 1.5 is the most exciting video model release from xAI. You can generate realistic video with synchronized audio in a single pass, capable of juggling complex motion with precise prompt adherence. We pushed it hard across a range of scenes, and came up with the ultimate prompting guide to get the most out of this model.

Video examples

Woman on the phone

This is a scene from a movie, she is having a conversation on the phone with a man, she speaks and pauses to listen then talks again. Her tone is hesitant then determined. She’s walking around slowly and talking on the phone. They are talking about a meetup on Thursday to discuss details.

Skateboarders

The guy in the middle starts skateboarding slowly, then goes faster, then is going very fast, the camera is following his face from the side, zoomed in on his face, he looks happy, documentary style coming of age vibe, the background buildings and street are moving fast as he’s skateboarding, handheld camera style, he’s laughing and says “I told you bro, this will be the best summer!”, then he goes silent, he’s looking at his friends briefly, and then he’s just looking at where he’s rolling with his skateboard, sounds of cars passing by, the two other teenagers are laughing, they are not talking. Skateboards rolling sounds.

Game character

The character is showing off her fighting moves, no music.

Skincare

She touches her cheek and smiles gently, then smiles wide looking into the camera, camera not moving, text and items stay exactly the same not moving.

Monkey

The monkey slowly turns a page, focused, reading. A small baby monkey dressed in a baby suit comes to distract him. No talking.

Architecture drone footage

Drone footage, flying at the front of the building.

Breaking wave

The wave crests fully and pitches forward, the translucent green crest folding and crashing down onto the dark rocks with tremendous force. White foam explodes upward and outward, hanging for a moment before collapsing back. Sea spray drifts across the frame in the dawn wind. The water rushes back off the rocks in white rivulets. A second smaller wave rises behind. Sound: the deep boom of a heavy swell hitting rock, the hiss and rush of water pulling back across stone, the low moan of wind across an open coastline, sea spray on a microphone.

Dawn run, Bangkok

The camera tracks alongside the runners in Bangkok as they continue sprinting, all four pumping their arms and legs in perfect lockstep, breath visible in the cool morning air. The lead runner glances briefly at the camera, then snaps his focus back ahead. The shopfront shutters, parked motorbikes, and pedestrians on the curb streak past in heavy horizontal motion blur. Sound: the rhythmic slap of running shoes on pavement, heavy synchronized breathing, the rumble of a distant scooter, the muffled chatter of an early morning street market.

Quiet afternoon

The figure slowly lowers the phone from their ear, exhales, and lets their hand fall to their side. They turn their head almost imperceptibly toward the room. Dust motes drift through the shaft of golden light. The cat lifts its head, ears swivelling. The CRT television flickers once. The sheer curtains stir in a slow breeze. Sound: distant city traffic muffled through glass, a kitchen tap dripping somewhere off-screen, the soft hum of the old TV, the creak of a wood floor.

Hopefully, these give a good sense of just how far you can push Grok Imagine Video 1.5.

How to prompt it

After experimenting with Grok Imagine 1.5 quite a bit, we came up with the following prompting tips that can really elevate your outputs.

Write the Sound: section like a sound designer

Every example above has an explicit Sound: section. Signaling this to the model and describe how you want sound to be designed in your video can make or break the final delivery.

Vague: Sound: city sounds, traffic.

Specific: Sound: cars passing by, skateboards rolling on pavement, teenagers laughing, the distant rumble of a street.

It knows the difference between ambient noise and targeted sound design. You can be as granular as you want, and it will keep up.

A few things that work particularly well: “sounds of cars passing by,” “skateboards rolling sounds,” “no music,” “camera not moving.” These are all spatial and material cues that tell the model exactly what the soundscape needs to feel like.

Use intensity modifiers

Without them, the model picks its own interpretation of scale. “The wave crests” is ambiguous. “The wave crests fully and pitches forward, crashing down with tremendous force” is much more indicative.

The skateboarding scene works, for instance, because of “starts skateboarding slowly, then goes faster, then is going very fast” and “background buildings and street are moving fast.” Remove those words and you get a flat, static clip.

Describe camera movement

The model holds static if you don’t ask for movement which is generally the right call if you don’t specify anything. A locked camera with patient motion reads more cinematic than unnecessary moves. But when you want a certain camera move, be sure to stipulate that.

Things that work: slow push-in, aerial push-in toward, camera drifts gently to the left, tracking shot alongside, locked, static. The Iceland clip asks for “slow aerial push-in” and “camera drifts gently to the left as it descends.”

Keep it focused

The model handles focused prompts better than sprawling ones. The skincare scene is three beats: touch cheek, smile gently, smile wide at camera. The monkey scene gives each character a clear role. You can really hone in on specific actions while keeping everything else locked and still.

Starting with the image

The best way to use Video 1.5 is to start with a still you’ve already dialed in. Use any image generator, like Grok Imagine Image, or your own photo to nail the composition and lighting first. Once the frame looks right, the video prompt only needs to say what changes.

Iridescent form

Starting image:

Abstract 3D render of a large glossy morphic form — smooth curved surfaces of transparent glass or liquid chrome, refracting prismatic iridescent color bands of cyan, magenta, gold, and electric blue against a pure black background. Hyperreal studio lighting, physically accurate reflections and refractions.

Abstract iridescent morphic glass form

Then passed to Video 1.5:

The glossy morphic form slowly undulates and breathes, its surfaces shifting like liquid mercury. The prismatic iridescent bands — cyan, magenta, gold, electric blue — flow and ripple across the curves as the shape subtly deforms and reforms. The light refracts differently as the surface tension shifts. The form rotates almost imperceptibly. Sound: a deep resonant hum, like the inside of a seashell, the faint crystalline ring of glass under tension, slow and meditative.

Wabi-sabi interior

Starting image:

A minimalist Belgian wabi-sabi interior. A long low linen-upholstered sofa in a sandy oatmeal tone sits against a tactile cream lime-plaster wall. A single rough-hewn dark walnut coffee table sits in front of it on a polished concrete floor. On a built-in concrete plinth: a squat ceramic table lamp with a dark earth-brown clay base and a soft cream linen shade, casting a low warm glow. A heavy linen throw drapes asymmetrically across the sofa. No decoration, no clutter, no pattern. The architectural photography of Vincent Van Duysen and Axel Vervoordt.

Minimalist Belgian wabi-sabi interior

Then passed to Video 1.5:

The afternoon sunlight coming through an unseen window slowly shifts and dims as time passes. The shaft of warm golden light that falls across the linen sofa and concrete floor moves gradually to the right and narrows, the color shifting from warm amber to cooler blue as the hour advances toward evening. The lamp’s warm glow becomes more pronounced as the room darkens. Shadows deepen in the corners. Sound: deep interior quiet, the barely audible ambient hum of the city outside, a building settling in the cooling air.

The still handles composition and color while the video prompt handles motion. Keeping them separate can make both easier to iterate on.

Run it on Replicate