1328 hand-tested AI prompts across image, video, and music — each with full context, variables, and a sample output. Content is authored in Traditional Chinese; the UI here is English.
15 seconds is the gold-standard length for short-form video ads. Sora 2 handles narrative instructions better than other models — you can give it a three-shot mini-story instead of just a single shot.
Sora 2 natively generates dialogue — a `Dialogue: "..."` field produces the voiceover directly; a lived-in, cluttered background …
Specifying three powder colors (crimson/cobalt/gold), backlighting to make the particle edges glow, and a pure black background turn a …
Breaking `Style / Audio / Duration` into explicit slots — a single Audio line lists ambient music, birdsong, breathing, and footsteps …
The `Camera movement:` field includes time-segmented beats (a dolly forward for the first 3 seconds, then a push-in to the hands) …
Sora 2 responds best to this kind of writing for 'physical realism' — using a dedicated `Physics:` section to call out surface tension, …
The key to image-to-video is 'only describe the change, don't redescribe what's already in the frame' — this prompt uses five sections, `Transformation/Li…
Write the audio as a bulleted `Audio requirements:` spec — the bag-impact sound 'synced to the motion,' breathing, …
The hardest part of a two-person dialogue is 'keeping both characters stable' — this prompt uses `[Character Cameo: …]` to separately lock in the persona of the host and the guest, …
Mood pieces most easily turn into a pile of adjectives — this one instead uses five fields, `Mood / Atmosphere / Detail…
Pairing a timelapse with metamorphosis hands t…
The night setting maxes out the contrast of or…
A translucent, light-transmitting ice wall plu…
It fully exploits native audio, roaring wind p…
FPV racing's biggest risk is losing the sense …
Flying alongside the eagle and matching its mo…
Generates a UGC-style spoken-word ad that looks like it was shot by an ordinary person — natural, unpolished, and well suited to short-form product-selling content.
A soothing macro slow-motion product close-up clip, well suited to beauty, food, and texture-driven products.
2026 is the year of image-to-video — use a single character reference image so the people in Sora 2 / Veo 3.1 videos stay consistent from start to finish, without drifting.
For e-commerce hero shots and ad spots. 8 seconds, seamless loop.
For B-roll, transitions, and intro filler. Pure atmosphere, no narrative.
A short-form talking-head clip — 8 seconds of an expert, teacher, or KOL speaking directly to camera. Easy to lip-sync in post when you dub in voiceover, and the delivery looks natural rather than stiff.
B-roll footage to intercut into documentaries, long-form YouTube videos, and brand story films — not the main shot, but atmospheric filler that supports the narrative. Sora 2 is particularly good at capturing mood in shots with no people in them.
For B2B client proposals, corporate anniversary and annual report videos, and similar corporate footage — the crucial first 3 seconds. Lets viewers immediately register that this is a company that takes its business seriously.