Image · Video · Audio — every model, one library

Skip the
blank page.
Paste the aha.

Free, curated prompts for AI image, video & music — tagged by model, tested by creators. Copy, paste, create.

1,026 curated prompts21 models coveredFree forever

Editors' picks

Featured this week

Create a 30-second Pixar-quality 3D animated miniature construction comedy where dozens of tiny construction workers build and decorate a giant cupcake. VISUAL STYLE: Premium cinematic 3D animation, adorable miniature world, bright pastel colors, warm golden lighting, ultra-detailed CGI, realistic food textures, shallow depth of field, smooth cinematic camera movement and satisfying visual details. Cheerful, playful and wholesome atmosphere. CHARACTERS: Tiny construction workers wearing colorful safety helmets and work uniforms. Keep their appearance and scale consistent throughout. 0–5s — ESTABLISHING Wide cinematic aerial dolly-in toward a giant cupcake sitting on a table. Dozens of tiny workers arrive with miniature cranes, forklifts, mixer trucks and construction vehicles. They quickly organize around the cupcake and begin preparing the site. Soft cheerful music, tiny engines, wheels and construction sounds. 5–10s — FROSTING Medium tracking shot as a miniature mixer truck pours silky pink frosting over the cupcake. Workers spread the frosting smoothly with tiny trowels, carefully covering the entire surface. Close-ups show the glossy frosting flowing and being leveled perfectly. ASMR: creamy pouring, soft scraping and tiny tool sounds. 10–17s — TOPPINGS Close-up cinematic montage. Miniature conveyor belts deliver fresh strawberries, blueberries, chocolate chunks and colorful rainbow sprinkles. Workers carefully position each topping with tiny tools and teamwork. Sprinkles fall in satisfying slow motion and sparkle under the warm lighting. Add subtle conveyor sounds, tiny footsteps and light magical chimes. 17–23s — CHERRY LIFT Low-angle cinematic shot of a huge shiny red cherry being lifted toward the cupcake by a miniature construction crane. Workers pull ropes and guide it carefully into position while the crane slowly rises. Build playful suspense with crane motor sounds and a gentle orchestral rise. 23–30s — FINAL REVEAL The cherry lands perfectly on top. The workers celebrate as colorful confetti explodes into the air. Some jump, others wave their helmets while tiny construction vehicles honk. Camera performs a smooth 360° cinematic orbit around the completed cupcake, revealing the beautifully decorated frosting and toppings. End on a satisfying hero shot with warm golden light, cheerful music, tiny cheers and a soft celebratory musical finish. No dialogue, no subtitles, no text. Maintain consistent character scale, realistic food physics, smooth animation, detailed miniature interactions and premium cinematic quality throughout.
FORMAT 10 seconds, 2 shots. Shot 1: Indoor bedroom, 0.0–5.0s. Shot 2: Outdoor night street, 5.0–10.0s. Transition: seamless hard cut exactly on shoe impact. Style: photorealistic live-action, premium cinematic commercial, realistic movement/physics, authentic handheld cinematography. SHOT 1 — INDOOR BEDROOM | 0.0–5.0s ENVIRONMENT Minimalist young woman's @[Image 1](image_1) private bedroom in daylight; clean, modern, intimate, naturally lived-in. Large curtain several meters behind Sharon, simple desk/console farther behind, minimal personal objects and subtle decoration. Enough depth for subject separation. Background softly out of focus; Sharon remains sharp and readable. CAMERA / LIGHTING Eye-level handheld, medium-wide/full-body framing with room for lateral movement and shoe actions, moderate shallow depth of field. Soft natural daylight, natural skin tones, soft highlights, gentle realistic shadows, warm clean atmosphere. Camera is operated physically with Sharon: lateral tracking, short push-in, reactive pull-back, quick pans, tilts and reframing. Never perfectly smooth or automated; movement is quick, light, imperfect and physically motivated by her actions. ACTION + CAMERA 0.0–1.1s — Sharon enters from RIGHT holding TWO SHOES, one in each hand, and walks toward center. Camera: handheld lateral tracking LEFT, physically stepping sideways while keeping medium-wide framing. 1.1–1.8s — Sharon @[Image 1](image_1) stops near center, raises one shoe toward camera as if showing it, then brings it toward her nose. Camera: short handheld push-in; keep face and shoe sharp. 1.8–2.4s — Sharon smells the shoe, immediately recoils with surprised disgust, pulls her face away and gives a short head shake. Camera: small reactive pull-back, then immediately settle toward her face. 2.4–3.0s — Sharon casually throws the first shoe toward the side of the room. Camera: quick reactive pan following the shoe briefly, then fast pan back to Sinta; believable handheld imperfection. 3.0–3.8s — Sharon shifts sideways, focuses on the second shoe, and throws it vertically upward. Camera: tilt up following the shoe; as it reaches the top of its trajectory, quickly tilt back down toward Sinta while keeping awareness of her and the descending shoe. 3.8–4.4s — Sharon tracks the descending shoe with her eyes, adjusts footing, rotates slightly sideways and prepares a high side kick. Camera: short lateral handheld adjustment keeping body and falling shoe in frame; subtle natural Dutch tilt may develop. 4.4–5.0s — Shoe descends into kicking range. Sharon performs a fast, precise HIGH SIDE KICK: leg extends sharply sideways, torso rotates slightly, supporting leg grounded, arms naturally counterbalance, eyes locked on the shoe. Full real-time. Camera: quick handheld lateral reposition and reactive pan following the kick; frame momentarily shifts with the force. EXACT MATCH CUT — 5.0s Hard cut at the exact moment Sharons extended foot contacts the airborne shoe. The kick is still at full speed and the cut happens DURING the action. Outdoor shot continues the same body direction, leg extension, torso rotation and momentum. SHOT 2 — NIGHT / QUIET CITY STREET | 5.0–10.0s ENVIRONMENT Relatively quiet modern urban street at night with broad paved roadway, streetlights, scattered storefronts, illuminated windows, trees, signage and distant buildings creating layered practical lighting. Sparse distant traffic. Spacious, cinematic, quiet. A WHITE SPORT MOTORCYCLE is behind Sinta and slightly toward one SIDE of frame, not centered directly behind her. Open street creates a clear diagonal path from Sharon to the motorcycle. LIGHTING / OUTFIT Naturalistic night lighting: warm streetlights mixed with cooler ambient city illumination, realistic highlights on black riding outfit, natural reflections on white motorcycle and asphalt, cinematic contrast with realistic exposure. Sharon wears a coordinated black-and-white motorcycle riding set: fitted black racing jacket with clean white shoulder/sleeve panels; fitted black short riding pants with subtle white detailing; compact black knee/shin protection; black motorcycle gloves with small white accents; clean white low-top sport sneakers; white crew socks to mid-calf; loose natural hair; no helmet. Motorcycle: white sport motorcycle. CAMERA Handheld kinetic cinematography. Operator physically follows Sharon using diagonal tracking, short arc movement, reactive panning, slight push-in and quick reframing. Never a perfectly programmed path; camera reacts to her changing body direction. ACTION + CAMERA 5.0–5.5s — HARD CUT. Sharon appears outdoors continuing the indoor side kick: kicking leg extended, torso rotated, momentum continuing. White sport motorcycle partially visible behind her toward one side. Camera immediately reacts. Camera: quick handheld pan/reframe catching the continuing kick trajectory. 5.5–6.3s — Sharon lowers her kicking leg naturally and does not stop. She moves DIAGONALLY across frame toward the motorcycle, forward-diagonal rather than straight backward/sideways, gradually turning torso and hips toward it while moving. Camera: handheld diagonal tracking with slight forward tracking plus lateral shift; keep Sharon dominant while motorcycle remains visible near the side. 6.3–7.2s — Sharon continues diagonally; body progressively rotates toward motorcycle. This is a traveling diagonal turn, NOT a stationary turn-in-place. Feet keep carrying her forward while torso rotates. Camera: short handheld arc around Sharon while continuing diagonal tracking; slight angle change and subtle pan to keep her centered; motorcycle remains partially visible at the side. 7.2–8.0s — SLOW MOTION BEGINS. Sharon is already mid-diagonal movement toward motorcycle. Body continues diagonal rotation; loose hair swings naturally; jacket moves w
subject_definitions: : The sole primary subject: the same young adult woman shown in the reference, with pale skin, long dark hair in an elaborate updo with pale ornaments, refined features, and a graceful fashion-model presence. She wears the reference pale silver-gray embroidered gown and translucent shawl, and expresses quiet happiness while trying on the outfit. Source(s): . summary: Create a 15.083-second vertical fashion scene in one unbroken observational take. A dark-haired woman in an ornate pale silver-gray gown lifts and drapes a translucent shawl, adjusts the outfit, admires herself in a mirror, and finishes in a graceful full-body reveal that communicates the joy of trying new clothes. Preserve the supplied reference’s identity, elegant interior, luminous palette, and detailed textiles. Hard constraints: Exactly one continuous vertical shot with no cuts or hidden transitions. | The same woman remains the sole primary subject throughout. | Show her trying on and arranging the translucent garment over the pale silver-gray embroidered dress, conveying genuine delight in the new outfit. | Preserve the reference identity, wardrobe, ornate bright interior, mirror, garment rack, luminous cool-blue and champagne palette, and fashion-editorial detail. | Keep the presentation tasteful and focused on clothing, posture, silhouette, and fabric movement. | No narration, subtitles, or spoken dialogue. Generation controls (authoritative) — Prompt-text rule: every section label and descriptive sentence is silent production direction, never speech content. Do not speak, recite, chant, sing, or read this prompt aloud. Background music: enabled. Generate a context-appropriate non-diegetic score unless the user explicitly requests silence or no music. Dialogue permission is enabled, but this plan contains no dialogue lines. Generate no character speech, conversation, chanting, singing, or other vocals. Dialogue language: automatic. Preserve each explicit language tag; otherwise use the language required by the scene and user intent. Narration: disabled. Generate no narrator, voice-over, spoken description, or audible reading of prompt text, section labels, instructions, asset tags, or metadata. Subtitles: disabled. Do not render subtitles or dialogue captions. Creative direction (authoritative within explicit user intent and reference fidelity) Priority: Apply these creative directives clearly wherever they do not conflict with explicit user requirements, identity retention, or fixed first/last-frame composition. Format and direction: Favor believable observation, motivated reframing, natural performance, and environmental detail over polished spectacle. Retain small imperfections that support authenticity while keeping the subject readable. Progress from a human-scale anchor to environmental scale, preserving depth layers and landing on a strong reveal composition. Camera execution: Follow the action continuously in a single unbroken take, using only motivated pans, small translations, and reframing needed to preserve spatial clarity. Look: Treat connected visual references as authoritative for identity, wardrobe, objects, environment, composition, lighting, color, and medium; do not impose an unrelated restyle. Pacing and composition: Use moderate, readable motion with natural acceleration and enough settling time for each important beat. Use exactly one continuous shot with no hidden or explicit cuts; express all action beats through blocking, performance, camera motion, reframing, and synchronized sound. An explicit user-requested shot or cut count remains authoritative over this density preference. Compose for a vertical frame: protect faces, hands, and essential action inside the central mobile-safe area; favor depth movement and avoid losing subjects at narrow side edges. retention_analysis: -> identity: Use the supplied collage as the authoritative visual reference for the same woman’s identity, facial features, pale skin, long dark hair in an elaborate updo with pale ornaments, downcast gaze, refined feminine silhouette, pale silver-gray embroidered gown, fitted bodice, flowing floor-length skirt, gold floral ornamentation, jeweled details, translucent wide-sleeved shawl, garment rack, bright ornate interior, tall windows, pale-blue mountain scenery, decorative furnishings, warm small lights, oval mirror, layered depth, luminous window backlight, cool silver-blue and warm champagne palette, detailed textiles, graceful portrait and near-full-body framing, and polished dreamlike fashion-editorial treatment. Preserve these elements consistently throughout one continuous scene, while treating the collage as a reference rather than reproducing its nine-panel layout. Use the depicted poses to guide the natural action of lifting the sheer garment, draping it around the shoulders, checking the waist and silhouette, and turning toward the mirror with quiet delight. Present the performance with believable observational imperfections and tasteful focus on clothing, posture, silhouette, and fabric movement.. detailed_description: [Shot 1] At 00:00.000-00:15.083, At the start, stands beside the pale garment rack in the luminous ornate interior, already wearing the pale silver-gray embroidered gown and holding the translucent shawl. She notices the garment with quiet anticipation, lifts it carefully, and brings it around her shoulders. Her expression changes into visible, restrained delight as the sheer fabric catches the window light. Without any cut, she smooths the shawl and checks the embroidered waist and flowing skirt, allowing the camera to follow her natural movement and preserve her graceful silhouette. She takes a few small steps toward the oval mirror, turns slightly to inspect the front and rear drape, then settles into a poised full-body stance with a soft satisfied smile and a brief downward glance at the finished o
Create a vertical 4:5 cinematic travel-diary poster built around one beautiful family moment in Santorini, Greece. TOP — REAL PHOTOGRAPHY: A young couple with their little daughter at a scenic Santorini viewpoint during golden-hour sunset. The daughter is naturally positioned between her parents while both parents gently kiss her cheeks at the same time. She closes her eyes and gives a genuine happy smile. Make the interaction spontaneous, affectionate and completely natural — like a real family vacation photograph, not models posing for an advertisement. Behind them: iconic white Santorini architecture, blue-domed churches, Mediterranean Sea, distant cliffs, small boats and vibrant bougainvillea. Warm sunset light wraps naturally around their faces and hair. Real skin pores, tiny imperfections, realistic hair strands, accurate hands, natural fabric texture and believable shadows. Shot like a premium full-frame travel photograph, 35mm lens, shallow but realistic depth of field, cinematic dynamic range. Absolutely photorealistic — no AI-looking skin, excessive HDR, plastic faces or artificial expressions. Keep generous negative space around the subjects for subtle editorial typography. Add only a few refined details such as: “TRAVEL DIARY” “Sept 11, 2026” “SANTORINI — GREECE” and one small handwritten travel note. BOTTOM — HAND-PRINTED MEMORY: Instead of simply duplicating the photograph, reinterpret the same family moment as an original vintage travel-print artwork. Use imperfect screen printing, risograph dots, engraved linework, faded ink, rough edges and authentic paper grain on warm ivory stock. Build the Santorini landscape around the family as a graphic illustration: simplified blue domes, cliffside houses, sea, sunset and bougainvillea integrated naturally into the composition. Limited ink palette of deep Mediterranean navy, sun-faded terracotta orange, warm ivory and tiny touches of dusty blue. Use a large expressive hand-painted title: “MORE GOOD DAYS” Surround it with only a few carefully placed diary elements: a Santorini postal stamp, tiny handwritten notes, one or two taped miniature travel photographs and “A SMALL DIARY — #001.” Keep everything intentionally imperfect and tactile rather than digitally clean. The transition between photography and illustration should feel like a torn page from a personal travel journal rather than a basic 50/50 split. Overall aesthetic: real family vacation photography × independent travel magazine × vintage European tourism poster × handmade screen print. Warm, intimate, nostalgic and premium. Strong enough to stop someone while scrolling, but never overcrowded. The photograph must feel genuinely captured in real life, while the lower artwork feels physically printed by hand. No Chinese text. No copied layouts. No generic AI collage aesthetic. No excessive stickers. No fantasy. No fake-looking faces. No malformed hands. No waxy skin. No over-saturation. Keep all three family members consistent between the photographic and illustrated sections.
0.00–2.50 Warm ivory background. Large black text: **O GORDAO DOS CORTES INOVOU.** Text slides in fast from the left with a clean ease-out. A thin red line grows underneath it. On the beat, the whole composition compresses slightly and snaps back. 2.50–5.50 The red line extends across the screen and becomes a crop transition. Reveal huge stacked text: **POSSANTE COMO UM JATO.** **VELOZ COMO UM FOGUETE** Use aggressive kinetic typography: large scale, edge cropping, fast horizontal movement, strong black/ivory/red composition. Keep the text sharp when it lands. 5.50–8.50 The typography collapses into five rectangular video-frame cards. The cards rapidly fill one after another while a red timing line underneath moves much slower. Use shallow 2.5D perspective, fast card stacking, whip movement and clean parallax to visually show that generation is finishing ahead of time. inspired in some brazilian podcast 8.50–11.00 The five cards snap together into one flat composition. Reveal: **O PRIMEIRO CORTE CHEGA.** **EM APENAS 15s.** Make **EM APENAS 1.5s** the biggest visual element. Use a fast mask reveal, strong scale change and one heavy motion-design impact when the full statement lands. 11.00–12.70 A red line shoots around the frame and builds a rectangular border. The previous text disappears through the moving crop. The border contracts several times in sync with an accelerating heartbeat-like bass pulse. 12.70–15.00 Brief silence. One final impact. Reveal: **JÁ TA DISPONÍVEL BIG.** **https://t.co/XaNKtgyaIs** Warm ivory background, black typography, thin red frame. Everything becomes completely flat and still. Hold unchanged until the end. STYLE Premium 2D motion design with very subtle 2.5D card depth. Swiss editorial typography mixed with energetic commercial motion graphics. Warm ivory, black and vermilion red. Large type, strong cropping, fast masks, cards, lines, parallax, precise easing, short motion blur only during transitions. No full 3D scenes, no UI, no particles, no random objects, no glitches, no letter-by-letter animation.
Create a 30-second Pixar-quality 3D animated ASMR comedy short aboard a warm, lantern-lit pirate ship galley. STYLE: Premium cinematic 3D animation, expressive characters, polished feature-film quality, warm amber lighting, realistic food textures, cinematic depth of field, playful slapstick humor, gentle ship movement and crisp immersive food ASMR. Characters: A large, burly, bearded pirate chef wearing a worn apron, and a mischievous bright-green parrot constantly trying to steal ingredients. Keep both characters visually consistent throughout. 0–4s Extreme close-up of the parrot stealing a garlic bulb. The pirate's hand suddenly slams beside it. They freeze and stare at each other. Brief record-scratch silence. The pirate flicks the parrot away; it spins through the air and lands on a pot rack pretending nothing happened. 4–8s The pirate rapidly chops garlic with crisp ASMR. The parrot tiptoes toward a tomato. Garlic hits hot oil with a huge sizzle, startling the parrot and sending it tumbling off the rack. Metal cups clang. 8–14s Fast cinematic cooking montage: tomatoes sizzling, herbs being torn, olive oil pouring in golden slow motion, sauce bubbling and the pirate confidently stirring the skillet. Layer detailed chopping, sizzling, pouring and bubbling ASMR. 14–19s The parrot spots the food and secretly tries to drag the skillet away. Its tiny body strains comically across the wooden floor. The pirate slowly turns around and stares. The parrot freezes while still holding the handle, then innocently whistles and lets go. 19–24s The pirate finishes cooking with a dramatic skillet toss. Steam rises as the glossy dish is plated. The parrot watches hungrily, trying to look innocent. 24–30s Instead of scolding the parrot, the pirate prepares a tiny plate of sauce and bread and slides it across the table. The parrot happily bounces and eats beside him. They exchange a satisfied look as the ship gently sways. End with a wide cinematic shot of the cozy galley glowing under amber lanterns, subtle ocean sounds and wooden ship creaks, with one soft playful accordion note fading out. No dialogue, no subtitles, no text. Maintain strong character consistency, natural physics, smooth animation, cinematic framing and detailed food ASMR throughout.
Create a curiosity-driven character vignette built from 10–15 short cuts around the character from Image1. Preserve the character’s identity, visual style and overall design. Infer its nature and world directly from Image1. Open with an immediately intriguing visual situation that creates a simple unanswered question. Build the following cuts around discovery, reaction, progression and small consequences. Each cut should introduce a new visual idea while clearly advancing the same underlying situation. Let the character reveal itself through behavior rather than exposition. Use its movement, instincts, habits, abilities, limitations, relationships and interaction with the surrounding world to make the sequence increasingly interesting. Keep the narrative simple enough to understand without dialogue. Create visual cause and effect between cuts. Allow details introduced early to gain meaning later. Build toward a clear reveal, reversal, transformation, emotional beat or satisfying visual payoff near the end. Avoid generic daily routines, disconnected montage imagery and repetitive actions. Do not force familiar human behavior onto the character or world. Let the situations emerge naturally from the reference. Make every few seconds visually distinct through changes in scale, framing, environment, movement, tension and information. Maintain strong continuity so the viewer wants to see what happens next. Use objective third-person cinematic observation with natural handheld imperfection and varied shot sizes. Keep the camera readable, responsive and visually motivated. Use only clean diegetic sound. No dialogue, narration, BGM, subtitles or title cards.
Duration: 30 seconds Aspect ratio: 3:4 vertical screen Live-action themed transformation short film, combining character selection cards, action match cuts, and comic photo collages. Fast-paced, playful, and lively. [Character Reference] Image 1 is the sole character identity reference. Throughout the entire film, it is always the same adult woman, with stable facial identity and body proportions. The opening uses the character card's original outfit and original hairstyle. The following four outfits and the hairstyles after completing each transformation are executed as described below. Before each complete look appears, the opening hairstyle is kept first, and the specified action is used to switch the hairstyle, accessories, and makeup. Do not bring the character card's background or multi-view layout into the video. [Full Film Structure] Four silhouette cards are offered for selection at the opening. The display order is fixed as: yellow → orange → green → pink. Each set follows: silhouette card flips into a colored character card → new outfit with the opening hairstyle → specified action triggers the complete transformation → holding the corresponding plush toy for display → three-photo comic collage. At the end, review the four looks in sequence. [0–3 seconds: Choose a character] Light gray studio background, the character turns from a side profile toward the camera. Both hands form a rectangular viewfinder frame in front of the chest with thumbs and index fingers. Four floating vertical cards unfold below the chest. From left to right on screen, strictly: orange, pink, green, yellow. In the center of each card are different cute black plush toy silhouettes. The character's right index finger taps toward the cards, then makes a selecting gesture of grabbing an invisible thin line and pulling it back toward the body. The yellow card on the far right is selected and enlarges forward. The card flips, the black silhouette is revealed as a colored yellow plush toy image, then slides away to reveal the character. The card is an independent animated layer; the fingers are not required to actually grip the card. [3–8 seconds: Yellow look, raising the camera triggers the transformation] Light blue background, accented with yellow stars and comic-style small radiating lines. The outfit is a yellow short-sleeved collared shirt, a dark blue bow tie, and a dark blue star pleated skirt. The character first keeps the opening hairstyle, holding a yellow instant camera with both hands, the yellow camera strap hanging naturally. The character raises the camera from the chest to eye level, covering part of the face, and makes a photo-taking motion. A clean match cut is made on the beat when the camera covers the face. The next shot becomes the completed look: hair neatly styled, with a yellow star hair clip added; in hand is switched to a small yellow plush toy. The prop change is completed through editing; the camera does not melt into the plush toy within the shot. The character brings the plush toy close to the side of the face, the other hand on the hip, tilting the head and smiling. At about the last 1 second, enter the yellow comic collage: first a thick white outline appears along the character's silhouette, then the cutout of the character is scaled and arranged into the poster. In the foreground is one large full-body photo, and in two rectangular frames behind are close-up photos. All three photos belong to the same character and the same yellow look. The character photos remain still; yellow halftone dots, diagonal frames, and small plush toy stickers may move slightly. The orange silhouette card pops in from the foreground, flips to reveal an orange fox plush toy, then slides away to enter the next look. [8–14 seconds: Orange look, back turned hair whip transformation] Cream yellow background, orange diagonal cut border. Orange tube top, orange-and-white multi-layered ruffled short skirt. First keep the opening hairstyle, the character lightly lifts the skirt hem, body swaying with the beat. The character turns her back to the camera, then quickly turns her head back, whipping her hair. A match switch is made at the instant the hair whips and the face is obscured. When returning to the front, the hairstyle has become two long black braids, paired with orange-and-white fox ear hair accessories and star decorations. Facial identity remains consistent, and the long braids sway naturally with the turning momentum. First briefly show the complete look and skirt hem, then cut to the pose holding an orange fox plush toy. The plush toy is brought close to the side of the face, the character closes her eyes and laughs softly. At about the last 1 second, enter the orange comic collage: in the foreground, a playful full-body cutout with one foot lifted, and behind it two close-up freeze-frame photos. Thick white outlines, orange halftone dots, and radiating lines form the layout. The character photos remain still, while the background graphics retain slight animation. The green silhouette card pops to the foreground, flips into a green plush toy image, then slides away. [14–19 seconds: Green look, crown toss transformation] Mint green background. Light green corset-style top, yellow lacing and off-shoulder decoration, light yellow sheer long skirt. The character first keeps the opening hairstyle, holding a small golden crown on the palm of her right hand, displaying it in front of the chest. The character gently tosses the crown upward, palm rising with the motion. Switch the look on the beat when the crown rises and the character turns sideways. In the next shot, the character is already wearing the same crown, and the hairstyle has become long, fluffy red-brown curls. The crown's position is clear and stable; two crowns do not appear, and the crown is not shown passing through the head. The character keeps the side profile, raises a hand to complete the follow-through, then turns her eyes to the camera. Then cut to the pose of holding red roses in one hand and raising a green plush toy close to the side of the face with the other. At about the last 1 second, enter the green comic collage: the foreground shows a static full-body cutout in the long skirt, with two close-up photos behind. Green borders, halftone dots, a small crown, and plush toy stickers maintain the comic style. The expressions and hair in the photos remain still, while the overall layer can slide in. The pink silhouette card pops out, flips into a pink plush toy image, then slides away. [19–24 seconds: Pink-blue look, racket swing transformation] Light purple background, pink diagonal cut shapes, with a few tennis line-art accents. Pink and mint blue patterned sleeveless top, light blue denim pleated skirt, pink belt. First keep the opening hairstyle, the character holds a white tennis racket with both hands and does one horizontal swing. Match-switch on the beat when the racket and arm quickly sweep across the front of the body. When the swing ends, the look has already changed to buns pinned up on both sides, paired with small bows and hair clips. The swing direction and body pose are continuous; do not repeat a second swing, and do not use the racket to cover the entire frame. The character completes the follow-through pose, tilts her head, and smiles at the camera. Then cut to a medium close-up holding a pink plush toy, the plush toy placed beside the face, the other hand naturally on the hip. At about the last 1 second, enter the pink comic collage: one foreground full-body cutout and two close-up photos behind. The three character photos remain frozen, with clear white outlines. Background color blocks, halftone dots, and plush toy stickers may move slightly. [24–30 seconds: Review of the four looks] Each look is about 1.5 seconds, hard-cutting on the beat in the order yellow, orange, green, pink-blue. The background color corresponds to that look, and the medium close-up composition is consistent. Yellow: one hand raises the yellow plush toy, the other hand's index finger clearly points at the plush toy, with a playful expression. Orange: one hand raises the orange plush toy, the other hand makes a small fist near the side of the face, winking once. Green: raise the green plush toy close to the side of the face, gently tilt the head, close the eyes, and smile softly. Pink-blue: both hands hold the pink plush toy in front of the chest near the chin, making a brief blinking expression. Finally hold this pose and end on the music's final beat. [Comic Collage and Cards] The cards include the reveal process from black silhouette to colored plush toy; it cannot be only color blocks covering. The comic posters for the four looks use a consistent layout: one full-body subject, two close-up photo frames, white outlines, corresponding-color halftone dots, and plush toy stickers. The character photos must be truly frozen; they cannot continue blinking, talking, or swaying hair. The background graphics and the overall photo layer can move. The multiple figures in the collage are photos of the same person, not multiple real people appearing in the studio at the same time. [Music and Sound] Light, upbeat pop music with clear drum beats, running consistently throughout the film. Raising the camera, whipping the hair, tossing the crown, and swinging the racket each land on clear musical downbeats. Photo freeze-frames are paired with light shutter sounds, and card flips are paired with short card-flip sounds. No character dialogue, no speaking mouth movements, and no dialogue subtitles added. [Visual Restrictions] Do not add brand logos, random advertising text, or garbled text; retain the character cards, color blocks, plush toy stickers, and comic decorations. Do not change the character's facial identity, and do not add new outfits. The color and shape of each plush toy remain stable. No fused fingers, extra limbs, or prop clipping. The transformations rely on action match cuts, without using body melting, vortices, or smoke obscuring.
2x2 grid, 9:16, 4 goat chess players who passed away [SUBJECT] :: single standalone vertical tribute poster, not a grid Create a heroic illustrated legacy poster for [SUBJECT], showing them as a larger-than-life cultural icon. The poster should combine a giant emotional portrait, multiple smaller action poses, milestone trophies/medals, symbolic locations, fan signs, handwritten quotes, and inspirational typography into one dense but readable composition. CORE FORMULA :: GIANT_PORTRAIT([SUBJECT]) + ACTION_ECHOES([career moments]) + ACHIEVEMENT_OBJECTS([trophies / medals / records]) + HOME_SYMBOLS([city / country / team / arena]) + FAN_MEMORY([signs / banners / handwritten notes]) + LEGACY_TEXT([values / slogans / future impact]) = premium inspirational athlete collage poster AI-INFERENCE RULES :: - infer the sport, uniform, colors, trophies, signature pose, and visual symbols of [SUBJECT] - infer 4–8 smaller action poses showing iconic movements or career moments - infer the most meaningful achievement object: trophy, medal, championship, record, title, crown, ball, apparatus, etc. - infer background landmarks: home city, national flag colors, stadium, court, arena, skyline, mountains, track, field, or stage - infer 5–10 short text fragments: slogan, values, stats, fan signs, handwritten quote, legacy phrase - make the subject feel inspirational, mythic, beloved, and larger than the sport - keep the poster as ONE finished image, not a 2x2 sheet STYLE :: luxury sports tribute poster ::5 maximal illustrated collage ::5 watercolor + gouache + pencil texture ::4 premium magazine cover energy ::4 heroic portrait realism mixed with painterly poster art ::5 large elegant typography ::4 scrapbook/fan-art emotional layering ::4 soft atmospheric glow and paper grain ::4 national/team color palette inferred from subject ::3 dense but readable hierarchy ::5 COMPOSITION :: top = bold inspirational headline or name upper center = huge close-up portrait looking upward / determined / visionary middle = smaller action poses wrapping around the portrait lower center = trophy, medal, ball, apparatus, or iconic achievement object bottom = crowd, stadium, city, flag, court, field, or homeland scene edges = flowers, paint splashes, smoke, stars, signatures, banners, handwritten notes TEXT LOGIC :: Use large readable title text: [SUBJECT NAME] [LEGACY PHRASE] Add short inferred phrases like: “MORE THAN A CHAMPION” “REDEFINING WHAT’S POSSIBLE” “THE BEAUTIFUL GAME LIVES ON” “FLY HIGHER” “COURAGE LOOKS GOOD ON YOU” “DREAMS HAVE NO BORDERS” NEGATIVE :: no 2x2 grid, no plain portrait, no empty background, no corporate infographic, no photorealistic Photoshop collage, no flat vector, no random unrelated objects, no illegible text, no generic athlete, no watermark
{ "prompt_details": { "title": "Big Tech, Small Scale - The Sound Stage", "target_platform": "Video AI Generation (Kling / Wan / Seedance)", "aspect_ratio": "16:9", "duration_seconds": 10, "fps": 30 }, "scene": { "concept": "A miniature 3-inch presenter demonstrating high-end wireless headphones at scale.", "environment": "Pristine minimalist creative studio desk, transitions from a macro hardware environment to a full-scale real-world tabletop context.", "subject": { "appearance": "3-inch-tall charismatic male presenter, stylish modern streetwear with neon cyan accents, energetic body language.", "action_sequence": [ "0-3s: Sprints confidently across the curved, matte-black anodized aluminum headband of oversized wireless headphones.", "3-6s: Reaches an enormous tactile knurled volume dial, grabs it with both hands, and forcefully spins it, triggering a visible translucent sonic shockwave.", "6-8s: Takes a dynamic running leap off the arch directly onto the ultra-plush memory foam leather ear cushion, landing like it's a luxury beanbag.", "8-10s: Looks directly at the high camera with a wide smile, flashing a quick thumbs-up as the macro scale shifts to reveal the true size." ] }, "cinematography": { "camera_lens": "Macro probe lens, shallow depth of field transitioning to deep focus.", "camera_motion": "Starts at an ultra-low floor-level tracking shot chasing directly behind the presenter's heels, whip-pans around the dial interaction, then performs a rapid crane-up pullback into a wide tabletop hero shot.", "framing": "Extreme close-up macro framing to wide contextual flat-lay showing the headphones next to an open laptop and a ceramic coffee cup on a white marble desk." }, "lighting": { "style": "High-end commercial tech aesthetic.", "key_light": "Soft diffused top-down studio softbox.", "accent_lights": "Edge-defining cyan and electric magenta rim lighting tracing the brushed metal chamfers and matte textures." }, "render_qualities": { "resolution": "8K UHD", "visual_style": "Photorealistic commercial, Octane render look, Unreal Engine 5 hyper-detail, physical micro-textures, no motion blur artifacts, crisp silhouette separation." }, "audio_design": { "sound_effects": [ "Crisp rhythmic miniature sneakers squeaking on matte aluminum finish.", "Heavy mechanical tactile gear-clicks as the knurled dial rotates.", "Visceral low-frequency bass impact whoosh on dial turn.", "Soft leathery 'puff' thud upon landing into the memory foam cushion." ], "music": "High-energy, bass-driven modern tech-pop groove with punchy percussion and upbeat syncopation." } } }
[Reference Image Definition] Image1 = Red-dressed ethnic-style cartoon girl | Sole character reference Image1 is the only and highest-priority reference for the character's identity, face, hairstyle, clothing, and art style throughout the entire video. Throughout the entire video, strictly maintain the girl from Image1: Recognizable facial features Facial feature proportions Black double braids Red-and-blue braided hair decorations Ethnic-style headpiece Earrings Red ethnic-style long dress Clothing patterns and fur-trimmed decorations Brown short boots Soft, delicate 2D anime illustration texture Cute, lively, with a slightly playful and mischievous vibe Prohibited: Photorealistic rendering 3D rendering Turning into someone else Changing hairstyle Changing clothing Changing art style Image2 = X profile page screenshot | Sole background and layout reference Image2 is the sole reference for the background, UI, typography, and composition of the entire video. Strictly maintain from Image2: Top banner image Position of the circular avatar frame White UI background All buttons Icons Page structure Whitespace Colors Proportions Throughout the entire video, the camera is completely fixed. Any of the following is prohibited for the entire page: Zooming Panning Rotation Perspective changes Recomposition Except for the specified performance, the UI and background must not change. [Overall Requirements] Generate a 15-second, 9:16 vertical, fixed-camera, 2D cartoon-animation-style short video. Throughout the entire video: The background is always the X profile page screenshot from Image2 Only the character and the text being manipulated move The character is initially located inside the circular avatar frame The character's appearance must be replaced with the red-dressed ethnic-style cartoon girl from Image1 The circular avatar frame position and the page itself remain unchanged at all times The overall rhythm is light, clear, with a touch of comedy No dialogue No narration No speech bubbles No additional subtitles No new unrelated characters No new unrelated logos No new unrelated text [Timeline Action Design] [0.0–2.3 seconds] The entire X page remains completely fixed. Only the inside of the circular avatar frame begins to move. The red-dressed girl inside the circular frame first blinks gently, revealing a sweet smile. She slightly lowers her head to adjust her cuffs, then gently lifts the edge of her skirt, as if preparing to make an entrance. Then she looks up at the page outside, glances down at the text content on the page, and finally looks back at the camera, a playful "up to mischief" smirk slowly spreading across her lips. Movements are small, natural, cute, and lively. No exaggerated distortion. No large swinging motions. [2.3–4.0 seconds] The girl places both hands on the edge of the circular avatar frame, treating it as a small entrance, and naturally climbs out from inside it. The sequence of actions is clearly shown as: Both hands grip the frame edge Upper body leans out first One foot steps out first The other foot follows She gently lifts her skirt slightly to prevent the long dress from getting caught The skirt hem briefly gets lightly caught on the frame edge She looks back and gently tugs at the skirt hem The skirt comes out smoothly She lands lightly on the white whitespace area below the X page Upon landing, the character should be light and natural, with the skirt, braids, earrings, and headpiece swaying gently and naturally. Only the character comes out of the avatar frame. The circular avatar frame itself remains in place, undeformed, and does not disappear. The page itself must not undergo any structural changes. [4.0–7.6 seconds] After landing, the girl looks around and notices there is a lot of text on the page. She immediately shows a delighted expression of "I've got a mischievous idea." She begins tearing off sections of text from the page one by one. The operable targets are all readable text / letters / numbers within the page, including: Profile name User ID Bio text Join date Following count Follower count Text inside buttons Post content text Other readable text / letters / numbers She tears off text only 3 times, with clear actions and crisp rhythm: 1st tear She walks to the area near the profile name, pinches an entire section of text with two fingers, and clearly peels it off the page surface like a sticker. The torn-off text is quickly folded twice in her hands and instantly becomes a small paper airplane. She flicks her wrist and sends the paper airplane flying toward the camera. The paper airplane passes through the foreground and flies out of frame. The corresponding original text position immediately becomes blank white. 2nd tear She tears off a section of bio text. This time the motion is more practiced — she folds it into a paper airplane quickly while walking. Then she crisply tosses it out toward the other side of the camera. The paper airplane flies across the foreground of the frame. The corresponding text disappears simultaneously, leaving only blank space. 3rd tear She tears off a larger block of text from the post area. This block is somewhat larger, and she quickly folds it into a slightly bigger paper airplane with both hands. She first does a small "ready to throw" motion, then suddenly throws it toward the camera. The paper airplane rapidly grows larger toward the camera, then flies out of frame. The corresponding original text area is completely transformed into blank white space. After the 3rd throw, she proudly claps her hands lightly. Important requirements Text must be lifted and torn off like paper stickers It cannot simply disappear Paper airplanes must be folded from the torn-off text Each time a paper airplane flies away, the corresponding original text position must genuinely become blank Paper airplanes should not be so large as to block the entire frame Only text may be removed The following elements must be preserved: Top banner illustration Circular avatar frame Image areas All graphic icons Button outlines Dividing lines Thumbnails Other non-text visual elements [7.6–10.4 seconds] She looks around the page and notices there is still a lot of residual text. Her proud smile slightly fades, revealing a hint of "how is there still this much" expression. She lightly places one hand on her hip, then produces a simple broom from behind her. Then she performs 3 brisk screen-clearing motions: 1st sweep A horizontal sweep across the upper part of the page. An entire row of residual text slides away like thin scraps of paper. 2nd sweep A diagonal sweep from the middle of the page downward. The remaining text continues to be gathered and pushed toward one side of the page. 3rd sweep A firm but not overly forceful sweep toward the lower right of the page. The last pile of residual text is swept entirely out of frame. After sweeping, she lightly taps the broom handle on the ground once, as if satisfied with the finish. Important requirements Only text touched by the broom may disappear Text must be gathered and swept away following the broom's motion It cannot suddenly disappear without reason The broom may only appear after the paper airplane actions are complete There is only one broom — it cannot multiply After sweeping, all old text on the page must be completely cleared. Only the following remain on the page: Background image UI graphic icons Button outlines Image elements Large areas of clean white whitespace [10.4–13.7 seconds] She leans the broom gently to one side, then looks down and takes out a thick black marker from her outfit. She pops the cap off with a "snap," turns around, and faces the large area of white whitespace that has been cleared. Then, using the thick black handwritten marker, she writes two lines in the large blank area from the center to the lower part of the page: This place belongs to Sun Keke now! Thank you all for liking it These two lines of text must be clearly, accurately, and completely displayed. No typos. No missing characters. No garbled text. No substitution with other sentences. No addition of any other new text. Text style requirements: Black Thick marker handwritten feel Written out stroke by stroke naturally Large enough Clearly legible Occupying a prominent position in the frame The writing process must clearly show the pen tip advancing — the characters are written out stroke by stroke, not appearing instantly. [13.7–15.0 seconds] After writing, she puts the marker cap back on. She takes a small step back and looks at her "masterpiece" with satisfaction. Then she holds the broom with one hand and the marker with the other, slightly turning to the side. She then looks back over her shoulder toward the camera, revealing a slightly smug, cute, and playful mischievous smile. Not an exaggerated laugh, but a gentle eye-squint, the happy smile after a successful prank. Finally, there is a static freeze frame for approximately 0.7 seconds. In the freeze frame, the following must be clearly visible simultaneously: The two lines of large text on the page: This place belongs to Sun Keke now! Thank you all for liking it The red-dressed girl looking back at the camera with a playful smile The broom The X page background with all text cleared

How it works

Three steps. No prompt engineering.

Find a prompt that matches your idea, copy it, and create — in whatever tool you already use.

  1. 01

    Browse & filter

    Narrow the library by media type, model, or category — images, video, music, or voice.

  2. 02

    Copy in one click

    Every prompt is paste-ready, parameters included. One click puts it on your clipboard.

  3. 03

    Paste & create

    Drop it into Google AI Studio, Hailuo AI, Suno — anywhere. Got a great result? Share it back.

Why AhaPrompt

Built for multi-model creators

One library that keeps up with the whole AI creative stack — so you don't have to.

01Every major model
Prompts for Seedance, MiniMax H3, GPT Image 2, Nano Banana, Seedream 5, and more — each labeled with the model it was written for.
02Three media types
Images, video, and audio in one place. Filter by media type and jump straight to your craft.
03One-click copy
Complete prompts, parameters included, on your clipboard in one click. Paste into any tool or API.
04Curated + community
Editor-curated collections alongside community submissions — every prompt is reviewed before it goes live.
05Four languages
The full library in English, Chinese, Japanese, and Korean — copy in your language, create in any.
06Free forever
Browse, copy, and use every public prompt at no cost. Great prompts should be accessible to everyone.

Good to know

Frequently asked questions

Everything you need to know about AhaPrompt and how to get the most out of the library.

Free forever — no account needed to browse

Your next aha is one paste away.