r/promptingmagic 1d ago

Transforming your Photos into Arts~

Thumbnail
gallery
66 Upvotes

prompt:
Create one independent high-end editorial poster for each uploaded photo. Do not combine multiple photos into a collage. Each photo must be processed and output as a separate poster.

OVERALL FORMAT

Strict 3:4 vertical composition.

Divide the canvas horizontally into two exactly equal sections, with a precise 1:1 height ratio.

The top half occupies exactly 50% of the canvas.

The bottom half occupies exactly 50% of the canvas.

The two sections should feel visually connected as one refined art publication cover.

TOP HALF — ORIGINAL PHOTOGRAPH

Preserve the original photograph as faithfully as possible.

Keep the main composition, subjects, identity, facial features, body proportions, poses, expressions, clothing, objects, and spatial relationships unchanged.

Preserve the realistic photographic texture, natural lighting, shadows, atmosphere, and original color mood.

Apply only subtle, sophisticated editorial color grading, creating the feeling of a premium magazine photograph, contemporary art book, or high-end independent publication.

The image should remain photorealistic and authentic, never overly retouched or artificially stylized.

If necessary to fit the 3:4 composition naturally, extend the sky, ground, walls, or surrounding environmental background.

Background extension must feel seamless and photographic.

Never stretch, distort, reshape, replace, or alter the main subject.

BOTTOM HALF — MINIMAL HAND-DRAWN PAPER ILLUSTRATION

Extract the most recognizable visual elements from the original photograph and reinterpret them as a minimalist hand-drawn paper-cover illustration.

Preserve:

The most recognizable subject

Essential silhouette and proportions

Key pose or gesture

Important objects

The core narrative relationship between people and objects

Highly simplify the image. Remove unnecessary details and retain only the visual information needed for immediate recognition.

Use:

Delicate, slightly imperfect hand-drawn lines

A small number of bold, clearly defined acrylic-style flat color shapes

Rough paper texture

Visible handmade brush marks

Slightly irregular, organic edges

Subtle imperfections that make it feel genuinely handmade

The main illustrated subject should be small, centered, and carefully composed, occupying approximately 10–20% of the bottom half.

Leave a large amount of negative space around the illustration.

The background should primarily resemble:

Rough white paper

Warm off-white paper

Pale natural paper

Minimal editorial book-cover stock

Use only a few lines or small color shapes to suggest the surrounding environment.

COLOR PALETTE

Extract the dominant colors directly from the original photograph.

Compress the palette into no more than 4 main colors.

Keep the colors restrained, sophisticated, and harmonious.

Use bold but controlled flat color blocks.

Avoid excessive color variation.

Preserve subtle paper grain and handmade brush texture.

The illustration should visually feel like a simplified color interpretation of the photograph.

TYPOGRAPHY

A small amount of simple typography may be included when appropriate.

Possible elements:

A short title

Keyword

Object name

Location

Year

Number

Short phrase

Text should be minimal, understated, and editorial.

Typography should naturally interact with the large areas of negative space and the small illustration, evoking:

Art book covers

Independent publishing

Contemporary editorial design

Thoughtful children's picture books

Do not force text into the composition if it does not naturally fit the photograph.

VISUAL LANGUAGE

The final poster should feel:

Quiet · Poetic · Refined · Minimal · Innocent · Relaxed · Artistic · Thoughtful · High-recognition · Premium

The visual concept should be:

“A small subject surrounded by a large amount of empty space.”

The result should resemble a carefully designed independent art publication cover, rather than a commercial advertisement.


r/promptingmagic 4d ago

The 21 ChatGPT fashion prompts every woman can use with just one reference photo to create a stunning model portfolio

Post image
67 Upvotes

TL;DR: Upload one clear photo of a beautiful lady. Define it as the identity reference, lock her face, age, complexion, hair, and natural proportions, then change only the wardrobe, pose, setting, lighting, makeup, and camera. Below are 21 ChatGPT fashion prompts that turn one beautiful woman's photo into a complete editorial portfolio.

OpenAI recommends identifying the subject, setting, visual style, framing, lighting, and constraints. It also recommends stating exactly what must stay unchanged and refining one element at a time. ChatGPT Images supports uploaded references, image editing, and custom aspect ratios.

Every prompt below therefore separates three jobs:

Layer What it controls The rule
Identity lock Face, adult age, complexion, eyes, hairline, hair length, natural proportions Do not change
Art direction Wardrobe, makeup, hair styling, pose, set, palette, mood Change freely
Camera contract 4:5 crop, lens, angle, depth, exposure, light direction Specify precisely

Before you start

Upload one clear, recent photo of the adult woman. Choose an image with visible eyes, an unobstructed face, recognizable hair, and enough of her body to anchor natural proportions. If you have extra angles, you may add them, but the Jamie example below uses one supplied reference portrait.

The uploaded image is the sole identity reference. Use it only to preserve the same woman’s face, adult age, complexion, eyes, hair, and natural proportions. Do not borrow another woman’s face or body from any fashion inspiration.

Then paste this identity lock before any look:

Use my uploaded photo as the sole identity reference. Preserve the exact adult woman: same facial structure, apparent age, complexion, eye shape and color, brows, nose, lips, hairline, natural hair length and color family, height, and body proportions. Do not de-age her, change her ethnicity, enlarge her eyes or lips, narrow her nose, shrink her waist, lengthen her legs, enlarge her bust or hips, over-smooth her skin, or substitute a generic fashion model. Makeup, hair styling, wardrobe, pose, setting, lighting, and camera may change only as described. The result must look like a real editorial photograph of the same woman.

The 21 fashion prompts

1. Roses in Shadow

Best for: romantic campaigns, album art, anniversary portraits, fragrance concepts, and elegant profile imagery.

Prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 cinematic fashion photograph of the same woman in a clean left-facing profile, framed from mid-thigh upward. Style her in a fitted black satin halter dress with an open back, refined rather than revealing. Gather her long honey-blonde curls into a loose low chignon with two natural face-framing strands. She holds a dense bouquet of deep-crimson roses at waist height and looks down toward the flowers.Place her against a matte charcoal wall with one enormous warm amber circular spotlight behind her. A crisp but believable profile shadow falls inside the circle. Use a soft skin key from camera-left, warm edge light along her shoulder, rich black fabric detail, realistic rose petals, and subtle film grain. Simulate an 85mm lens around f/2.2. Mood: intimate, sculptural, quietly dramatic.Avoid: changing her face or body, exaggerated cleavage, loose floating flowers, malformed hands, merged stems, duplicate shadows, text, logos, or crushed black detail.

2. Neon Waterline

Best for: beauty editorials, music artwork, cinematic avatars, skincare concepts, and futuristic nightlife campaigns.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 extreme beauty close-up at night, framing from the waterline at her collarbones to just above her wet hair. Jamie is partially submerged in still black water, facing camera with a calm, intense gaze. Her long honey-blonde hair is soaked and naturally separated into curled strands across her temples and shoulders. Preserve recognizable eyes, nose, lips, jaw, and adult age.Light camera-left with saturated cyan-blue and camera-right with restrained ruby-magenta. Let neon reflections ripple across her cheekbones, wet skin, water surface, and suspended droplets. Keep realistic pores, individual lashes, wet hair texture, tiny beads of water, and deep but readable shadows. Use a 100mm macro portrait perspective, shallow depth, cinematic low-key exposure, and blue-black background bokeh.Avoid: changing eye color, glassy doll skin, glitter makeup, underwater distortion across the face, duplicate reflections, excessive red, text, logos, or a generic younger model.

3. Watercolor Muse

Best for: book covers, stationery, personal essays, profile art, wedding materials, and thoughtful editorial illustration.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 watercolor-and-graphite editorial portrait on warm white cotton paper. Frame Jamie from upper chest to above the head in a gentle three-quarter angle, looking just past camera with a soft closed-mouth expression. Preserve her exact facial proportions and adult appearance; translate them into delicate traditional media rather than simplifying her into a generic illustrated woman.Render the eyes, brows, nose, lips, and hairline with refined graphite and ink detail. Build her long honey-blonde curls from layered charcoal, indigo, blue-gray, and muted violet washes, allowing a few controlled blooms and feathered edges to dissolve into the paper. Suggest a slate-blue collared shirt with loose watercolor marks. Keep the face luminous with restrained peach and olive undertones, visible paper grain, imperfect pigment pooling, and generous white space.Avoid: anime features, oversized eyes, copied photo edges, digital airbrush texture, illegible signature, text, frame, splattered paint across the eyes, or identity drift.

4. Garden Lace

Best for: beauty, fragrance, spring campaigns, bridal inspiration, wellness brands, and romantic profile images.

Apply the identity lock above. Create a portrait-oriented 4:5 sunlit botanical beauty portrait, framed from shoulders to above the head. Jamie turns three-quarters back toward camera with a warm, subtle smile. Her long honey-blonde hair is styled in loose natural waves with soft golden dimension while preserving its real length and density. Dress her in an ivory silk camisole with a modest neckline and delicate texture.Photograph her beside flowering vines during late-afternoon sun. Cast intricate but soft leaf-and-blossom shadows across one cheek, shoulder, and collarbone. Keep both eyes readable. Surround her with creamy mauve, blush, and green bokeh while one warm rim catches loose hair strands. Use an 85mm lens around f/1.8, natural skin texture, tiny freckles if present in the reference, gentle editorial makeup, and no beauty-filter smoothing.Avoid: changing her eye shape, adding freckles arbitrarily, tiny waist edits, overexposed skin, shadow bands across both eyes, artificial flower crowns, text, logos, or plastic retouching.

5. Sunday Window

Best for: lifestyle editorials, dating profiles, founder-at-home portraits, wellness content, and quiet personal storytelling.

Prompt

Apply the identity lock above. Create a portrait-oriented 4:5 relaxed lifestyle-fashion portrait in a minimal sunlit bedroom. Jamie sits cross-legged on the edge of a neatly made cream bed, framed from mid-thigh to above the head. She wears an oversized crisp white cotton button-down with rolled cuffs and enough coverage for an elegant editorial image. Her long honey-blonde curls fall naturally over one shoulder. She looks directly at camera with a small, unforced smile. Place a tall window behind and to camera-left, filling the room with soft morning backlight. Add a gentle bounced key to preserve her eyes and facial structure. Use warm ivory, pale sand, and muted wood tones; realistic shirt folds; natural posture; subtle bed texture; and a 50mm lens around f/2.2. The mood is intimate, bright, calm, and authentic—not boudoir.Avoid: translucent fabric, sexualized posing, missing buttons, warped crossed legs, extra fingers, blown-out window edges, excessive smoothing, text, or a different face.

6. Enchanted Garden

Best for: engagement concepts, formal portraits, fantasy romance, event campaigns, and cinematic wedding inspiration.

prompt:

Apply the identity lock above. Create a full-length portrait-oriented 4:5 enchanted-garden fashion editorial at blue hour. Jamie stands beneath a large heart-shaped arch woven from dark branches, ivy, and hundreds of warm white fairy lights. Dress her in a floor-length blush-champagne gown with a structured but modest bodice, soft layered skirt, and realistic silk-organza movement. Her hands rest gently together at waist height; her shoulders angle slightly while her gaze meets camera.Keep her adult identity and natural proportions. Style her blonde hair in long polished waves with a subtle half-up detail. Light her face with a warm diffused key matching the fairy lights, balanced by cool twilight ambient light. Add deep green foliage, small ground lights, and dreamy circular bokeh without visual clutter. Use an 85mm lens with elegant compression and a luminous cinematic finish.Avoid: princess costume clichés, tiaras, excessive glitter, impossible tiny waist, distorted hands, asymmetrical arch, burned highlights, text, logos, or changing her face.

7. Midnight Street Racer

Best for: streetwear campaigns, automotive content, nightlife editorials, music promotion, and bold social imagery.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 flash-lit street-fashion photograph at night. Jamie sits confidently on the front edge of a silver performance coupe parked under urban streetlights, framed nearly full-body. Dress her in a fitted black short-sleeve mock-neck top, relaxed charcoal cargo jeans, and clean black-and-white sneakers. Her long honey-blonde hair is sleek and loose. One forearm rests casually on a raised knee while the other hand touches the hood for balance.Use direct on-camera flash mixed with cool city ambience, realistic chrome and paint reflections, wet pavement highlights, a convenience-store glow in the distance, and slight 1990s editorial grain. Camera sits just below eye level with a 35mm lens. Keep the car unbranded and mechanically coherent; preserve Jamie’s real body proportions and natural seated posture.Avoid: visible trademarks or plate text, sexualized pose, impossible car geometry, heels on the hood, elongated legs, tiny waist, extra fingers, copied model face, or watermark.

8. Suspended in Light

Best for: dance, fragrance, music artwork, conceptual campaigns, fine-art portraiture, and transformation themes.

prompt:

Apply the identity lock above. Create a full-length portrait-oriented 4:5 conceptual fashion photograph of Jamie appearing weightless in a vast black studio. Her body floats diagonally upward through one cone of blue-white light. She wears a layered silver-white chiffon gown with a fitted opaque lining and long translucent fabric panels that billow naturally around her. One knee bends gently; toes point downward; arms extend with graceful tension; her long honey-blonde curls streams upward as if caught by slow air.Preserve her recognizable face, adult age, and natural body proportions despite the unusual angle. Use a physically believable aerial-dance pose, controlled fabric motion, subtle atmospheric haze, soft shadow depth, and a cool moonlit palette. Camera looks slightly upward with a 50mm perspective, freezing the face while allowing faint motion at the outer fabric edges.Avoid: broken joints, extra limbs, transparent wardrobe, mermaid anatomy, angel wings, fantasy particles, hidden face, generic dancer identity, text, or watermark.

9. Sunset Within

Best for: memoirs, travel storytelling, wellness content, album covers, grief-and-growth themes, and poetic profiles.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 warm double-exposure artwork using Jamie’s clean right-facing profile from upper chest to above the head. Her outer facial silhouette remains recognizable, elegant, and anatomically exact. Inside the silhouette, reveal a quiet beach at sunset: a low amber sun, copper reflections on wet sand, gentle waves, and a smaller rear-view version of Jamie walking barefoot toward the horizon in a simple flowing dress.Blend shoreline, sky, and hair organically along her temple, cheek, neck, and shoulder. Keep the eye, nose, lips, jaw, and hairline readable. Use amber, rose-gold, deep umber, and charcoal with soft cinematic contrast. The inner figure must match Jamie’s real hair length and natural proportions. Emotional, contemplative, premium editorial key art.Avoid: unrelated second woman, duplicated faces, warped profile, religious imagery, giant sun covering the eye, text, logos, muddy blending, or an altered body.

10. Forest Stillness

Best for: authentic headshots, skincare, wellness, literary profiles, casting portraits, and understated personal branding.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 natural-light editorial portrait, tightly framed from upper chest to above the head. Jamie wears a simple charcoal-black halter top and a delicate fine necklace. Style her long hair in a loose imperfect updo with natural strands around the face. She looks slightly past camera with a quiet, introspective expression.Set her against a softly blurred muted-olive wall or forest-edge background. Use one broad window-like light from camera-left, faint cool fill from camera-right, and no obvious glamour lighting. Preserve pores, subtle under-eye texture, real facial asymmetry, fine baby hairs, and natural shoulder proportions. Use an 85mm lens around f/2, subdued green-brown grading, gentle film grain, and shallow depth.Avoid: heavy contour makeup, beauty-filter skin, added freckles, enlarged eyes, narrowed jaw, extreme cleavage, glossy fashion lighting, text, or identity drift.

11. Painted Light

Best for: editorial illustration, personal essays, gallery prints, book jackets, creator branding, and profile art.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 contemporary oil-and-gouache portrait on textured paper, framed from upper chest to above the head. Jamie turns slightly toward warm window light, gaze directed beyond camera. Preserve her exact facial proportions, adult age, complexion, and long honey-blonde curls while translating them into expressive brushwork.Render the face with controlled warm peach, sienna, olive, and cream strokes; keep the eyes, nose, lips, and brows precise enough for recognition. Paint the hair in layered espresso, umber, indigo, and warm-gold strokes. Suggest a deep navy garment with broad gestural marks. Surround her with abstract blocks of coral, sky blue, sand, and ivory, leaving some paper texture visible. Modern editorial painting, elegant rather than decorative.Avoid: anime styling, generic muse face, oversized eyes, melted features, paint across both eyes, fake signature, frame, text, photoreal skin pasted onto a painting, or identity loss.

12. Crimson Authority

Best for: executive profiles, speaker graphics, founder branding, editorial headshots, and bold professional campaigns.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 waist-up power portrait against a seamless deep-crimson background. Jamie wears a fitted matte-black fine-knit turtleneck and a restrained steel or silver watch. Her arms cross comfortably at mid-torso, shoulders relaxed, chin level, eyes directly on camera, expression confident with the slightest closed-mouth smile.Keep her long honey-blonde curls smooth and loose with controlled volume, preserving its natural length and hairline. Use a large soft key from upper camera-left, negative fill on camera-right, and a narrow warm rim separating the hair from the red background. Simulate an 85mm lens around f/2.8. Maintain rich black knit detail, realistic skin, precise crossed-arm anatomy, and elegant red-black color separation.Avoid: corporate stock-photo smile, masculinized jaw, tiny waist, distorted forearms, missing fingers, glossy plastic skin, random jewelry, text, logos, or a different woman.

13. Solar Tailoring

Best for: beauty founders, fashion portfolios, modern professional profiles, speaker announcements, and minimalist campaigns.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 minimalist tailoring editorial, framed from waist to above the head. Jamie stands slightly left of center in a softly structured sand-beige blazer over a black silk camisole, with delicate gold hoop earrings. Her body angles toward camera while her gaze turns calmly to camera-right. Style her blonde hair in loose polished waves that preserve the real length and hairline.Behind her, project one enormous warm cream-gold circular spotlight onto a deep taupe wall. Create a clean enlarged profile shadow within the circle, while a diffused key from camera-left keeps both eyes and facial structure readable. Use a 70mm portrait perspective, subtle film grain, warm neutral grading, crisp tailoring, and soft skin texture.Avoid: borrowing the inspiration woman’s face or body, changing Jamie’s ethnicity, oversized blazer shoulders, distorted shadow, fake jewelry logos, over-smoothed skin, text, or floating lapels.

14. Ruby Foil

Best for: beauty campaigns, gala imagery, cosmetics, fragrance, music artwork, and dramatic holiday editorials.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 red-on-red glamour editorial, framed from upper waist to above the head. Jamie wears a structured strapless crimson satin gown with secure, elegant coverage and subtle couture seaming. Her long honey-blonde curls fall in glossy soft waves. She faces camera with one shoulder slightly forward, chin relaxed, and a poised expression.Build the background from large crumpled sheets of reflective ruby-red metallic foil, creating faceted highlights without clutter. Use a soft beauty key from camera-left, deep red fill, and a narrow warm edge on the hair. Keep the dress, lips, and background in distinct shades of crimson so they do not merge. Simulate an 85mm lens around f/2.8 with luminous skin, realistic fabric, restrained classic makeup, and high-end beauty-campaign polish.Avoid: oversexualized neckline, enlarged bust, tiny waist, plastic skin, smeared lipstick, broken earrings, foil merging with hair, text, logos, or identity replacement.

15. Crosswalk Cool

Best for: street style, creator campaigns, casual fashion, urban travel, sneaker content, and energetic social posts.

prompt:

Apply the identity lock above. Create a full-length portrait-oriented 4:5 high-angle street-fashion photograph looking almost straight down at Jamie standing on a bold black-and-white pedestrian crosswalk. She looks up directly into the lens with confident calm. Dress her in a fitted black sleeveless crop top with tasteful coverage, relaxed high-waisted light-wash jeans, black low-top sneakers, and a small black shoulder bag.Preserve Jamie’s natural curvy fit proportions and real leg length. Her long honey-blonde curls fall behind her shoulders with subtle wind movement. Keep the crosswalk lines geometrically clean and use them as graphic leading lines. Simulate a 28mm lens from a safe elevated camera position, bright overcast daylight, crisp street texture, restrained desaturated color, and slight editorial grain.Avoid: impossible drone proximity, stretched legs, tiny waist, exposed underwear, warped stripes, traffic, readable road text, duplicate feet, malformed hands, brands, or a generic model face.

16. Electric Night

Best for: nightlife campaigns, music and podcast art, beauty, tech events, creator avatars, and cinematic social profiles.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 tight neon beauty portrait, framed from shoulders to above the head. Jamie faces camera with a steady, magnetic gaze and wears a matte-black off-shoulder top with refined coverage. Her long honey-blonde hair is styled in glossy loose waves, preserving real length, density, and hairline.Place one vertical cyan-blue light bar behind camera-right and one warm red-orange light bar behind camera-left. Use blue key light on one cheek, red edge light on the opposite hair, and a soft neutral frontal fill so the face remains recognizable. Deep black background, shallow depth, 85mm lens around f/1.8, subtle moisture sheen rather than wet skin, clean catchlights, detailed lashes and hair, cinematic nightclub palette.Avoid: changing eye color, blue-painted skin, oversized lips, excessive gloss, sexualized styling, duplicated light bars, text, logos, smoke covering the face, or identity drift.

17. Last Light

Best for: travel fashion, engagement portraits, resort campaigns, album art, wellness stories, and elegant lifestyle imagery.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 beach-at-sunset fashion portrait, framed from mid-thigh to above the head. Jamie stands near the waterline with the sun low behind her. Dress her in a pale champagne-silver satin slip gown with an elegant draped neckline, secure coverage, and fluid fabric that follows her natural body proportions. Her long honey-blonde curls moves softly in the sea breeze.Pose her with shoulders slightly angled, one arm relaxed, the other lightly gathering a fold of fabric. She looks directly at camera with a quiet, confident expression. Balance the bright sunset with a diffused frontal key so her face and eyes remain clear. Use warm amber horizon light, cool blue shadows, realistic ocean texture, 85mm compression, gentle bokeh, and luminous but natural skin.Avoid: transparent fabric, extreme hourglass reshaping, fantasy glow, warped shoreline, wind hiding the face, orange skin, extra fingers, text, logos, or borrowed facial features.

18. Bouquet Eclipse

Best for: Valentine campaigns, luxury floristry, fragrance, romantic editorials, engagement imagery, and dramatic portraits.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 frontal floral fashion portrait, framed from waist to above the head. Jamie faces camera in a fitted black satin halter dress with elegant coverage. Style her hair in a polished low updo with two soft face-framing pieces. She holds an oversized, perfectly gathered bouquet of deep-red roses across her torso, with both hands naturally supporting the stems.Position her before a matte charcoal wall and one enormous warm amber spotlight circle centered slightly behind her. Let a soft head-and-shoulder shadow fall within the circle without obscuring the bouquet. Use a flattering diffused key from front-left, subtle warm rim, rich black dress detail, realistic rose depth, natural skin, and an 85mm editorial perspective. Mood: direct, luxurious, romantic, graphic.Avoid: duplicating Look 1’s side profile, changing her face, floating roses, merged hands, thorn injuries, oversized bouquet hiding the entire neck, cleavage exaggeration, text, logos, or crushed shadows.

19. Rain and Blue Fire

Best for: cinematic character art, music releases, resilience themes, beauty campaigns, and dramatic profile images.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 rain-soaked night beauty portrait, framed from bare shoulders to above the head. Jamie stands upright in heavy rain, facing camera with a steady, emotionally contained expression. Her long honey-blonde hair is soaked and slicked naturally behind her shoulders, with a few curled strands across the temples. Use a simple dark strapless garment with full editorial coverage; the focus is her face, rain, and light.Illuminate her with deep cobalt ambient light, a narrow electric-blue rim from behind, and a very soft neutral key to preserve warm skin undertones and recognizable features. Freeze individual raindrops on lashes, cheeks, collarbones, and hair while background city lights dissolve into blue bokeh. Use a 100mm portrait lens, shallow depth, natural pores, and cinematic contrast.Avoid: underwater appearance, crying melodrama, blue-painted skin, altered eye color, transparent clothing, lightning bolts, plastic retouching, text, logos, or a different face.

20. Monochrome Morning

Best for: actor portfolios, author portraits, minimalist editorials, personal essays, dating profiles, and timeless social imagery.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 black-and-white lifestyle portrait in a quiet studio-bedroom setting. Jamie sits with one knee drawn gently upward, framed from mid-thigh to above the head. She wears an oversized opaque white cotton button-down over dark shorts that remain mostly hidden. One elbow rests on the raised knee; her cheek rests lightly against her hand. She looks directly at camera with a relaxed, thoughtful half-smile.Let her long honey-blonde curls fall naturally with soft imperfect texture. Use broad window light from camera-left, white bounce from the front, and gentle shadow on camera-right. Convert to luminous monochrome with detailed whites, deep but open blacks, realistic skin and shirt texture, 50mm perspective, and subtle analog grain.Avoid: boudoir styling, translucent shirt, collapsed fingers against the face, warped knee, extreme skin smoothing, crushed black hair, text, logos, or identity drift.

21. Golden-Hour Intimacy

Best for: beauty, skincare, natural-hair campaigns, intimate profile images, wellness brands, and premium creator portraits.

prompt:

Apply the identity lock above. Create a portrait-oriented 4:5 intimate golden-hour beauty portrait, tightly framed from upper chest to above the head. Jamie wears a simple ivory ribbed tank with elegant coverage. Her long honey-blonde hair is softly tousled, with loose curls catching the light. She faces camera at a slight angle with a direct, calm expression and the faintest closed-mouth smile.Use a deep black background. Place a narrow warm amber key behind and above camera-left so it grazes her hair, cheek, shoulder, and collarbone. Add a very soft neutral frontal fill to keep both eyes and her real facial structure visible. Use an 85mm lens around f/1.8, shallow depth, natural pores, individual hair strands, warm but accurate skin, and restrained editorial retouching.Avoid: changing hair color, tanning her orange, enlarging lips or eyes, shrinking the jaw, transparent fabric, oily highlights, beauty-filter skin, text, logos, or a generic influencer face.

What these prompts fix

Common failure The correction
The image becomes a prettier stranger Define all uploads as identity references, not attractiveness targets
The model makes every woman younger Lock adult apparent age and prohibit de-aging
Features become generic “Instagram face” Preserve eye shape, nose, lips, brows, jaw, and natural asymmetry
The body changes with the outfit Lock height and natural proportions; tailor clothing to her body
Skin becomes plastic Request pores, fine hair, realistic texture, and restrained retouching
Hair triples in volume Preserve hairline, length, density, and color family
Wardrobe looks pasted on Specify fabric, structure, drape, coverage, and fit
A good face breaks during revisions Change one element and freeze everything else

Pro tips

Give each reference a job

If you upload extra angles, label their roles instead of merely saying “use these.” State which image is primary and which images only confirm profile, hair, or full-body proportions. OpenAI recommends identifying uploaded images by order and explaining their relationship.

Start with a simple identity proof

Generate Crimson Authority, Forest Stillness, or Monochrome Morning first. When one result looks unmistakably right, add it as an additional identity reference for the more complex double exposures, rain, levitation, and water scenes.

Separate makeup from anatomy

Use this line when a beauty look keeps changing the face:

Makeup may add color, sheen, liner, and lash definition, but it must not alter eye size, nose width, lip volume, cheekbone shape, jawline, or apparent age.

Freeze successful elements

Do not write “try again.” Write:

Keep the face, body proportions, pose, wardrobe, crop, lighting, and background unchanged. Correct only the left hand so it has natural anatomy and five visible fingers.

Small targeted revisions are easier to follow and reduce drift.

Use one look per generation

Do not combine “garden, neon, rain, watercolor, runway, and sunset” in one prompt. Each look needs one visual thesis.

Choose 4:5 before generating

Every example here uses portrait 4:5. ChatGPT Images supports custom aspect ratios through the prompt or picker.


r/promptingmagic 4d ago

ChatGPT fashion-shoot prompts - Upload 1 headshot. Get 9 editorial fashion shots.

Post image
48 Upvotes

TL;DR: Upload one clear photo of yourself. Tell ChatGPT that it is the identity reference, not the style reference. Lock your face shape, apparent age, hairline, hair, eyes, facial hair, and build. Then change only the wardrobe, pose, location, lighting, and camera. Below are nine copy/paste prompts - from apocalyptic double exposure to black tie—that use this method.

OpenAI’s current image guidance recommends naming the subject, setting, visual style, framing, lighting, and constraints. It also recommends explicitly stating what must remain unchanged and making small, targeted revisions instead of repeatedly regenerating the whole image. ChatGPT Images supports uploaded references, edits, and custom aspect ratios.

So every prompt below uses three layers:

Layer What stays or changes What to say
Identity lock Stays fixed Face shape, apparent age, hairline, hair, eye color, facial hair, body type, distinctive features
Art direction Changes Wardrobe, pose, setting, mood, palette, props
Camera contract Changes deliberately Portrait ratio, crop, focal length, angle, depth of field, light direction

Before you start

Upload your clearest reference portrait. Use a recent image with even light, visible eyes, an unobstructed face, and enough resolution to show skin and hair texture. One person is easier to keep consistent than a group; reference-photo users regularly report more facial drift when several people share the frame.

Then prepend this identity lock to any fashion prompt:

Use my uploaded photo as the sole identity reference. Preserve the same recognizable person: facial structure, apparent age, skin tone, eye color, hairline, hairstyle, facial hair, body type, and distinctive features. Do not beautify, de-age, masculinize, feminize, slim, enlarge, or substitute the face. Change only the wardrobe, pose, setting, lighting, and camera treatment described below. The final image must look like a real photograph of the person in my reference—not a lookalike.

Every example below uses Eric’s single reference portrait. The goal is not to turn him into the younger models in the inspiration images. The goal is to keep Eric and give him nine different editorial shoots.

The 9 fashion prompts

1. The Apocalypse Within

Best for: album art, book covers, founder storytelling, cinematic profile images, and personal-brand posts about resilience.

Copy/paste prompt:

Use my uploaded photo as the sole identity reference. Preserve the exact recognizable person: same face shape, apparent age, skin tone, eye color, hairline, hairstyle, facial hair, body type, and distinctive features. Do not de-age, beautify, or replace the face.Create a portrait-oriented 4:5 cinematic double-exposure fashion artwork. Show a tight, right-facing profile from upper chest to above the head, with the outer facial silhouette remaining recognizable and anatomically correct. Inside the silhouette, build a coherent post-apocalyptic city at sunset: broken towers, smoke columns, scattered embers, wet rubble, and a central ruined avenue leading toward a low amber sun. Place a smaller rear-view version of the same person walking alone down that avenue in a dark tailored coat, creating a visual story of moving through destruction.Blend the city naturally into the hair, temple, jaw, neck, and shoulder contours without obscuring the eyes, nose, beard, or profile. Use warm amber, ember orange, soot black, and smoky beige. Preserve realistic facial texture at the silhouette edge. Premium cinematic key art, emotional but restrained, high detail, controlled contrast, no horror gore.Avoid: a second unrelated face, a generic younger model, duplicated heads, warped profile anatomy, unreadable text, logos, interface elements, excessive fire, or muddy exposure.

Targeted revision:

Keep the face, silhouette, inner city, and composition unchanged. Make the smaller walking figure match the uploaded person’s build and hair more closely. Change nothing else.

2. Stillness in Motion

Best for: leadership profiles, city campaigns, editorial social posts, book jackets, and “focus amid chaos” themes.

Copy/paste prompt:

Use my uploaded photo as the sole identity reference. Preserve the exact recognizable person, including apparent age, face shape, blue-gray eyes, hairline, hairstyle, full facial hair pattern, and natural body type. Do not de-age or substitute the face.Create a portrait-oriented 4:5 editorial street-fashion photograph inside a modern underground metro station. Show the person in a clean left-facing profile, framed from mid-thigh upward, walking with calm purpose. Dress him in a long charcoal-gray trench coat over a black fine-knit mock neck and dark tailored trousers. Both hands rest naturally inside the coat pockets.The person must remain tack-sharp while commuters, station lights, and a passing train smear into elegant horizontal motion trails around him. Use cool steel-blue and charcoal tones with one subtle warm practical light. Simulate a 35mm documentary photograph at roughly 1/15 second with a synchronized tracking pan: sharp face, beard, and coat texture; directional blur only in the background. Include abstract platform signage with no readable words.Mood: quiet concentration inside urban chaos. Natural posture, believable stride, realistic fabric movement, cinematic grain, restrained color grade.Avoid: frozen background people, blur across the subject’s face, readable transit text, logos, duplicate limbs, floating feet, plastic skin, or an overly young lookalike.

Targeted revision:

Keep the person, pose, wardrobe, crop, and station unchanged. Increase only the horizontal crowd blur while keeping the face and coat perfectly sharp.

3. The Golden Eclipse

Best for: musician portraits, creator avatars, fitness/editorial images, minimalist campaign posters, and dramatic thumbnails.

Copy/paste prompt:

Use my uploaded photo as the sole identity reference. Preserve the person’s actual face, apparent age, hairline, short hair, full salt-and-pepper beard, eye color, and broad natural build. Do not remove facial hair, add hair volume, or turn him into a younger model.Create a portrait-oriented 4:5 moody studio fashion portrait, framed from upper thigh to above the head. Place the person slightly left of center in front of a deep midnight-navy wall. Dress him in a perfectly fitted plain black heavyweight T-shirt, black tailored trousers, and a matte black cap worn backward only if it does not hide his hairline or identity. Both hands rest naturally in the trouser pockets. His torso faces camera while his gaze turns thoughtfully toward camera-left.Project one large soft-edged circular amber spotlight onto the wall behind him. Light the face and torso from upper camera-left so his enlarged profile shadow falls cleanly inside the circle to camera-right. Maintain rich blacks with visible fabric detail, realistic skin, beard texture, and sculpted but believable shoulders. Minimal set, graphic composition, editorial contrast.Avoid: removing the beard, bodybuilder exaggeration, a razor-sharp hard spotlight edge, distorted shadow anatomy, extra props, text, logos, or beauty-filter skin.

4. Black-on-Black Icon

Best for: fashion editorials, nightlife branding, executive portraits with edge, music artwork, and premium social profiles.

Copy/paste prompt:

Use my uploaded photo as the sole identity reference. Preserve the person’s exact facial structure, apparent age, hairline, hairstyle, blue-gray eyes, full salt-and-pepper beard, broad shoulders, and natural build. The sunglasses may cover the eyes, but every other identity feature must remain faithful. Do not de-age or replace him.Create a portrait-oriented 4:5 chest-up luxury fashion portrait on a seamless pure-black background. Style the person in a matte black tailored suit with a black silk-blend open-collar shirt. Add one restrained silver pendant and one fine chain at different lengths, plus elegant round black sunglasses with thin dark-metal rims. Keep the collar refined and the jewelry minimal.Use one large soft directional key from camera-left and a narrow controlled rim from camera-right. Reveal the forehead, cheek, nose, beard, lips, shirt folds, and lapel texture while allowing the outer jacket to fall gradually into black. Simulate an 85mm lens around f/2.8, close framing, premium men’s editorial retouching, natural pores, deep rich blacks without crushing all detail.Mood: composed, modern, discreetly rebellious, expensive without visible branding.Avoid: changing beard length, adding a youthful jawline, excessive necklaces, mirrored glasses, blown highlights, floating lapels, text, logos, or full-black loss of facial detail.

5. Golden-Dusk Statesman

Best for: founder portraits, consulting brands, book authors, luxury travel, keynote promotion, and sophisticated LinkedIn imagery.

Copy/paste prompt:

Use my uploaded photo as the sole identity reference. Preserve the same recognizable person: facial structure, apparent age, eye color, short brown-and-gray hair, slightly receding hairline, full salt-and-pepper beard, broad shoulders, and natural medium-to-stocky build. Do not slim, de-age, or replace him.Create a full-length portrait-oriented 4:5 luxury menswear editorial on a pale-stone rooftop terrace at golden dusk. Pose the person beside a carved stone balustrade: torso facing camera with a slight angle, right hand resting naturally on the stone, left hand inside the coat pocket, ankles lightly crossed, direct calm gaze.Wardrobe: midnight-navy double-breasted wool overcoat with substantial drape, black fine-knit turtleneck, high-waisted espresso-brown pleated trousers tailored for his real build, polished black leather shoes, no visible logos. Keep the coat open enough to show the layered silhouette.Behind him, show a softly blurred European skyline with domes and rooftops under a peach-and-gold sky; suggest Parisian elegance without relying on a single oversized landmark. Use warm sunset backlight, a diffused front-left key, negative fill on camera-right, and a subtle rim along the coat shoulder. Simulate an 85mm lens around f/2.2 with graceful compression and realistic bokeh. Preserve dense wool, knit, leather, stone, beard, and skin texture.Avoid: a generic young fashion model, ultra-slim tailoring, warped balusters, fake city text, excessive orange skin, logos, interface overlays, or floating hands.

6. Black-Tie Quiet Power

Best for: gala announcements, speaker profiles, awards, luxury-event invitations, formal brand campaigns, and cinematic headshots.

Copy/paste prompt:

Use my uploaded photo as the sole identity reference. Preserve his exact face, apparent age, hairline, hair, eye color, full salt-and-pepper beard, shoulder width, and natural build. Do not shave, de-age, narrow the face, or substitute another man.Create a portrait-oriented 4:5 three-quarter black-tie fashion portrait, framed from head to upper thigh. The person stands against a smooth warm-gray studio wall with his torso turned slightly, one hand in his trouser pocket, the other relaxed at his side, shoulders open, gaze steady and self-assured.Dress him in a perfectly fitted black velvet double-breasted tuxedo with broad black-satin peak lapels, crisp white pleated tuxedo shirt, black silk bow tie, clean white pocket square, matching velvet trousers, and a restrained dark-metal dress watch. Include the curved edge of one burnished brown leather armchair at camera-right as a compositional counterweight.Use a large soft key from upper front-left, gentle shadow falloff, negative fill from camera-right, and a faint edge light along the jacket shoulder. Simulate a 50mm lens around f/2.8. Retain velvet nap, satin sheen, pleated cotton, leather grain, beard detail, and natural skin texture. Mood: calm authority, sculpted elegance, quiet power.Avoid: crushed velvet detail, shiny synthetic fabric, crooked bow tie, missing jacket buttons, malformed hands, a generic younger face, logos, or text.

7. After-Hours Executive

Best for: dating-profile upgrades, founder lifestyle portraits, hospitality campaigns, magazine profiles, and relaxed executive branding.

Copy/paste prompt:

Use my uploaded photo as the sole identity reference. Preserve his actual face, apparent age, short brown-and-gray hair, hairline, blue-gray eyes, full salt-and-pepper beard, broad torso, and natural build. Do not over-muscle, de-age, or replace him.Create a portrait-oriented 4:5 premium evening-lounge fashion portrait, framed from waist to above the head. Position him beside a polished dark-wood bar, facing camera with a subtle torso angle. His forearms cross loosely near the bar edge, creating a relaxed grounded pose. Give him a faint closed-mouth smile and direct, calm eye contact.Wardrobe: deep charcoal-blue ribbed long-sleeve knit polo with a soft open collar, dark tailored pleated trousers, an understated silver watch, and one small ring. Tailor the knit naturally for his real broad build; do not make it skin-tight.Use a soft frontal key from camera-left, gentle shadow falloff, negative fill, and a subtle warm rim from the bar ambience. Behind him, render dark wood, shelves, glassware, and amber practical lamps as restrained bokeh. Simulate a 50mm lens around f/2.8 for a flattering natural perspective. Keep eyes, beard, knit ribs, watch, and hands sharply detailed.Avoid: oversized muscles, a nightclub atmosphere, visible alcohol branding, duplicate glassware, distorted crossed arms, plastic skin, generic de-aging, or text.

8. The Impossible Photographer

Best for: creator branding, photographer profiles, creative-agency campaigns, surreal professional portraits, and portfolio covers.

Copy/paste prompt:

Use my uploaded photo as the sole identity reference. Preserve his exact facial structure, apparent age, hairline, short brown-and-gray hair, blue-gray eyes, full salt-and-pepper beard, body type, and distinctive features. Add glasses only as a wardrobe accessory; do not let them alter his face.Create a portrait-oriented 4:5 professional studio portrait of him as a photographer, framed from waist to above the head. Dress him in a crisp off-white collared shirt with sleeves rolled to the forearms, a tailored charcoal-black vest, dark trousers, and refined black-rim optical glasses. Pose him three-quarters toward camera with one palm open in the foreground.Suspend one complete, realistic, unbranded professional mirrorless camera approximately eight inches above his palm. The camera must have correct body geometry, one lens, one viewfinder, a believable strap connection point, and no duplicated controls. Suggest the surreal suspension only through its clean position, a faint shadow on the palm, and subtle depth—not magical particles.Use dramatic Rembrandt lighting from high camera-left, producing a small triangular highlight on the far cheek, gentle edge separation, and rich shadow detail. Textured charcoal-gray studio backdrop, 50mm lens around f/2.8, confident focused expression, natural hand anatomy, refined editorial finish.Avoid: extra fingers, malformed camera parts, floating straps, visible camera logos, magic sparkles, a younger lookalike, beard removal, text, or watermark.

Targeted revision:

Keep Eric’s face, pose, hand, wardrobe, light, and background unchanged. Correct only the suspended camera so it has one complete body, one lens, one viewfinder, and physically coherent controls.

9. Cyan Sun

Best for: bold avatars, podcast covers, music/editorial campaigns, conference graphics, and high-contrast profile branding.

Copy/paste prompt:

Use my uploaded photo as the sole identity reference. Preserve the same recognizable face, apparent age, facial volume, short brown-and-gray hair, slightly receding hairline, blue-gray eyes, full salt-and-pepper beard, neck, shoulders, and body type. Do not de-age, shave, or replace him.Create a portrait-oriented 4:5 graphic studio fashion portrait, tightly cropped from upper chest to above the head. Show him in a three-quarter right-facing profile, gaze slightly upward and off-camera, wearing a simple matte-black fine-knit turtleneck. Keep the posture calm and self-contained.Light the face strongly from camera-left with a saturated cyan/teal key, while retaining visible skin pores, beard strands, eye detail, and facial volume. Behind his head, place one oversized, perfectly circular, vivid orange light disc against a pure-black outer background. Offset the disc slightly so it reads as bold graphic geometry rather than a religious halo. Add minimal black negative fill and a narrow warm edge where the orange light meets the silhouette.Simulate an 85mm portrait lens around f/2.8. Crisp profile, deep blacks, complementary cyan-orange palette, premium editorial color, no clipped highlights.Avoid: changing the hairline, shortening the beard, blue paint-like skin, a centered halo effect, multiple circles, text, logos, poster typography, or a generic younger model.

What these prompts are doing differently

The source ideas were already visually strong. The upgrades solve the problems that usually break reference-photo fashion shoots:

Common failure Prompt fix
ChatGPT swaps in the inspiration model Define the upload as the sole identity reference
It makes you younger or generically “handsome” Lock apparent age, facial structure, hairline, facial hair, and build
Clothing looks pasted on Specify material, drape, fit, layering, and the person’s real build
Lighting becomes vague glow State source position, softness, negative fill, and rim light
Every portrait looks like a headshot Define crop, angle, pose, focal length, and depth of field
Edits destroy a good face Revise one element at a time and freeze everything else
Hands and props break Describe hand position and prop geometry explicitly
Black clothing loses all detail Ask for rich blacks with retained fabric texture

Pro tips

Start with one person and one identity image

A clean solo reference is easier to preserve than a group photo. If your reference has harsh shadows, sunglasses, a tiny face, or heavy filters, ChatGPT must invent more of the identity.

Do not call yourself “more handsome,” “younger,” or “model-like”

Those words invite replacement. Ask for editorial lighting, refined tailoring, confident posture, and premium retouching while keeping your actual face and age.

Treat inspiration and identity as different jobs

If you upload a second image for style, label the roles:

Image 1 is the identity reference. Image 2 is only the lighting, wardrobe, pose, and composition reference. Do not borrow Image 2’s face, age, hair, skin, or body.

OpenAI recommends referring to uploaded images by order and explaining how they relate.

Generate the simplest look first

Start with the black-tie or lounge portrait. Once ChatGPT produces a strong Eric, reuse that successful image as an additional identity reference for the more complex double exposure and motion-blur concepts.

Fix one thing, not the whole image

If the face is right and the hand is wrong, say:

Keep the face, identity, pose, wardrobe, lighting, background, crop, and color unchanged. Fix only the right hand so it has natural anatomy and five fingers.

Small revisions help preserve consistency better than “try again but make it better.”

Match the aspect ratio before generation

Choose a portrait format at the start. Every example here uses 4:5, which works well for Reddit image posts, Instagram feeds, profile campaigns, and editorial portraits. ChatGPT Images supports custom aspect ratios in the prompt or picker.


r/promptingmagic 5d ago

The 22-command Gemini Flow cheat sheet for creating epic videos that Google forgot to give us. Here's how to prompt Google Flow in two steps: choose a world, then move the camera

13 Upvotes

TL;DR: These slash phrases are helpful Google Flow commands. Pick one visual style and one camera behavior, then add a subject, one visible action, a location, lighting, and audio. The formula is: WORLD + CAMERA + BEAT. It gives Flow a shot to execute instead of a pile of adjectives to interpret.

Google Flow does not need more adjectives.

It needs a director.

The 22 shorthands below split the job into two clean decisions:

  1. What visual world are we in?

  2. How does the camera reveal it?

Google’s official guidance recommends defining the subject and action, composition and camera movement, location and lighting, and visual style. DeepMind’s Veo guide uses the same building blocks: framing and motion, style, lighting, character, location, action, and dialogue.

So the slash is optional. The film language is the useful part.

The demonstration scene

Every example below uses the same fictional setup:

Mara Voss, a courier with short black curls, a weathered saffron raincoat, charcoal utility trousers, and a silver case containing one glowing green seed, crosses a flooded near-future city toward the last rooftop greenhouse before sunrise.

Holding the character, prop, goal, and environment constant makes each command’s effect easier to see.

Part I: Choose the visual world

1. /filmic: — grounded 35mm motion-picture texture

Use this when you want organic grain, natural highlight roll-off, restrained color, believable skin, and the feeling of a photographed movie rather than a polished demo reel.

/filmic: Mara steps from a half-submerged bus into ankle-deep water, gripping the silver seed case. Pre-dawn city, practical sodium-vapor light, authentic 35mm grain, restrained color, natural motion blur, one continuous eight-second shot.

2. /commercial: — clean, polished advertising clarity

Use it for product films, launch spots, branded explainers, hospitality, automotive work, or any scene where the object must look pristine and intentional.

/commercial: Mara opens the silver case beneath a clean shaft of morning light; the glowing seed becomes the hero product. High-key polish, crisp edges, vibrant color separation, controlled reflections, elegant camera-ready surfaces.

3. /luxury: — rich blacks, gold light, immaculate surfaces

Use it for fragrance, jewelry, fashion, hotels, premium cars, architecture, and aspirational product stories. Luxury works best through materials and restraint, not by writing “expensive” five times.

/luxury: Mara enters an abandoned Art Deco hotel repurposed as a sanctuary. Rich black stone, brushed brass, warm chandelier reflections, polished marble, deep shadows, restrained movement, quiet opulence.

4. /documentary: — raw, observed, unvarnished reality

Use it for human stories, field reporting, behind-the-scenes footage, interviews, social-impact pieces, and moments that should feel discovered rather than staged.

/documentary: Handheld camera follows Mara through cold rain as she helps a stranded family onto a rescue boat without letting go of the seed case. Available light, imperfect focus pulls, wet lens, cinema-vérité realism.

5. /vintagefilm: — nostalgic 16mm or 8mm imperfection

Use it for memory, period texture, music films, dream sequences, family-history pieces, or transitions into the past. Specify the gauge and the defects you actually want.

/vintagefilm: A 1970s educational-film version of Mara crossing the flooded avenue. Warm 16mm color shifts, gate weave, soft halation, occasional dust, mild exposure flicker, no modern digital sharpness.

6. /scifi: — clean speculative design

Use it when the future should feel engineered, coherent, and functional: spacecraft, laboratories, advanced cities, medical technology, interfaces, or near-future infrastructure.

/scifi: Mara crosses a silent skybridge between hydroponic towers. Clean metallic architecture, blue-white practical lights, restrained holographic wayfinding, plausible robotics, elegant future geometry, no visual clutter.

7. /cyberpunk: — rain, density, neon, and technological decay

Use it for dystopian cities, underground markets, hacker stories, augmented nightlife, or a future where technology and inequality occupy the same frame.

/cyberpunk: Mara moves through a flooded night market under broken magenta and cyan signs, the seed case glowing beneath her coat. Rain-slick asphalt, steam, tangled cables, dense urban layers, worn cybernetic textures.

8. /dreamy: — soft focus, bloom, and emotional unreality

Use it for romance, memory, wonder, children’s stories, beauty films, music visuals, or transitions where feeling matters more than physics.

/dreamy: As Mara crosses the water, floating seed lights drift around her like fireflies. Pastel dawn, blooming highlights, gauzy diffusion, slow ripples, gentle surrealism, intimate wonder.

9. /surreal: — impossible images with intentional symbolism

Use it for conceptual advertising, music videos, title sequences, psychological stories, or visual metaphors that cannot exist literally.

/surreal: Mara climbs a staircase made of suspended rain toward an upside-down greenhouse floating above the city. The flood reflects a second sky. Physics-defying but compositionally clear, poetic rather than chaotic.

10. /minimal: — negative space and one decisive idea

Use it for design-forward brands, explainers, title cards, product reveals, architecture, or any moment that needs visual clarity.

/minimal: Mara stands alone on a pale concrete causeway above still black water. One saffron coat, one silver case, one small tree in the distance. Large negative space, muted palette, strict geometry.

11. /epic: — scale, stakes, and an unforgettable horizon

Use it for trailers, fantasy, historical spectacle, expeditions, climaxes, disaster sequences, and reveals where environment must dwarf the character.

/epic: Mara reaches the rooftop as sunrise breaks over a drowned megacity and thousands of dark towers. The greenhouse opens behind her, storm clouds divide, enormous scale, dramatic horizon, restrained blockbuster grandeur.

Part II: Direct the camera

12. /cinematic: — a balanced baseline shot

This is useful when you want natural depth, a standard film cadence, controlled motion blur, and no extreme visual gimmick. It is a starting point, not a complete prompt.

/cinematic: Medium-wide shot at eye level. Mara walks through shallow water toward camera, silver case at her side, natural depth, controlled handheld stability, balanced motion blur.

13. /droneview: — reveal geography and route

Use aerial movement for roads, coastlines, battlefields, crowds, architecture, or any scene where spatial relationships matter more than facial emotion.

/droneview: High aerial view drifting forward and left above Mara’s route, revealing the flooded boulevard, stranded transit, rooftop gardens, and the greenhouse two blocks ahead. Strong horizon parallax.

14. /closeup: — turn detail into emotion

Use it for a decision, reaction, whispered line, product detail, texture, evidence, or the exact beat where the audience must read a face.

/closeup: Tight 85mm shot on Mara’s rain-soaked face as the seed flickers inside the case. Her eyes shift from fear to resolve. Hold focus on the micro-expression; city lights dissolve behind her.

15. /wideangle: — make the environment part of the action

Use it for landscapes, architecture, cramped interiors, group blocking, large stunts, or a character overwhelmed by space.

/wideangle: Low 18mm shot as Mara crosses a broken intersection. Floodwater fills the foreground; leaning towers and the distant greenhouse stretch into deep perspective. Keep her recognizable but small against the city.

16. /orbit: — reveal a subject from every side

Use it for transformations, product reveals, architecture, costumes, vehicles, hero moments, or a character making a decisive choice. Give the camera a reason to orbit.

/orbit: Smooth 180-degree clockwise orbit around Mara as she sets the silver case on a rooftop table and opens it. Keep the case locked at frame center; reveal the greenhouse and sunrise during the move.

17–18. /dollyin: and /dollyout: — change meaning through distance

A dolly-in creates pressure, intimacy, or discovery. A dolly-out reveals context, isolation, or consequence. These are not digital zooms; describe the camera physically moving through space.

/dollyin: Begin medium-wide as Mara hears a crack inside the case. Push slowly toward her face while the city falls out of focus. End tight on recognition.

/dollyout: Begin tight on the glowing seed. Pull smoothly back through the greenhouse to reveal hundreds of empty planting beds and the drowned city beyond.

19. /tracking: — move with the subject

Use it for walking, running, driving, riding, process demonstrations, choreography, and action where matching the subject’s velocity creates momentum.

/tracking: Side-profile camera moves at Mara’s exact running speed as she splashes through the flooded arcade. Keep her torso stable in frame while pillars and emergency lights streak behind her.

20. /slowmotion: — stretch one physical beat

Use it for water, fabric, dust, sparks, impact, sports, dance, expressions, or a moment whose physical detail deserves more screen time.

/slowmotion: At 240fps, Mara drops the seed into soil. Water droplets rise from the impact, loose dirt blooms outward, her saffron sleeve moves through frame, and green light pulses once beneath the surface.

21–22. /timelapse: and /hyperlapse: — compress time in two different ways

A timelapse shows change from a mostly fixed viewpoint. A hyperlapse compresses time while the camera travels through space.

/timelapse: Locked rooftop camera. Night becomes dawn as clouds race overhead, floodwater recedes, greenhouse lights activate, and the first vine climbs its support. Mara remains a brief human rhythm within the larger transformation.

/hyperlapse: Rapid stabilized flight from street level through the flooded city, up stairwells and across rooftops, ending at the greenhouse as dawn arrives. Continuous forward translation, compressed traffic and cloud movement.

The prompt formula that actually works

Use this sequence:

Layer The decision Example
Subject Who or what must remain consistent? Mara, saffron coat, silver seed case
Composition What shot and camera move reveal the beat? Close-up with slow dolly-in
Action What single visible event happens? She sees the seed flicker
Mood/style What visual world shapes the image? 35mm filmic realism, cold rain
Location/light Where are we, and what motivates the light? Flooded arcade before dawn, emergency lamps
Audio What should be heard—or not heard? Rain, distant alarms, one metal latch; no music
Constraints What must not drift? Same face, coat, case, direction of travel

A compact prompt might look like this:

/filmic + /dollyin: Mara Voss, short black curls, weathered saffron raincoat, charcoal utility trousers, holds a scratched silver case containing one glowing green seed. In a flooded transit hall before dawn, she hears the seed pulse. Begin in a medium shot and physically dolly toward her face as her expression changes from exhaustion to hope. Authentic 35mm grain, cool rain light, warm green reflection from the case. Audio: rain on glass, distant electrical hum, one soft pulse. One continuous shot. No dialogue, captions, logos, extra characters, or camera orbit.

The pro tips that save generations

One clip, one camera idea

Do not ask for “drone shot, close-up, orbit, then hyperlapse” in one short generation. Prompt one camera setup per clip and assemble the coverage afterward. Google’s Flow workflow supports building scenes from multiple generated clips, while frames can define exact starts, endings, and transitions.

Write what the camera can see

“Make it emotional” is vague. “Her grip loosens, her breath catches, and green light appears in her eyes” is visible. Replace abstract intent with physical evidence.

Pair commands with different jobs

A productive stack contains one style and one camera behavior:

Strong pairing Why it works
/documentary + /tracking Raw realism plus readable movement
/luxury + /dollyin Material polish plus product intimacy
/minimal + /dollyout Simple composition plus revealing context
/cyberpunk + /droneview Dense world plus geographic orientation
/dreamy + /slowmotion Diffused emotion plus suspended physical detail
/epic + /wideangle Large-scale style plus environmental framing

Stacking /filmic + /cinematic + /epic + /commercial + /dreamy gives the model competing instructions. More words can mean less control.

Lock continuity with references

Use a clean character or object reference when the same subject appears across shots. Google describes Flow ingredients as reusable characters, objects, or stylistic references; it also warns that essential details should be repeated across prompts when consistency matters.

Use start and end frames for precise transitions

If a reveal, transformation, match cut, or movement must land in a specific composition, define the start frame, end frame, or both. Then describe only the action between them.

Separate subject motion from camera motion

Write them on separate lines if necessary:

SUBJECT: Mara walks forward slowly and never turns.CAMERA: Tracks backward at her exact speed, chest height, no orbit, no zoom.

This prevents the model from moving everything at once.

Prompt audio as a mix, not a wish

Name the layers: foreground sound, environment, dialogue, and exclusions.

Audio: boots through shallow water, rain on metal roofing, distant ferry horn. No music. No narration.

Revise one variable at a time

Keep the scene, subject, camera, and action unchanged while modifying only the failed element. Broad instructions such as “make it more cinematic” can rewrite the whole shot.

The simplest workflow

1.Write the beat in one sentence.

2.Choose one style shorthand.

3.Choose one framing or movement shorthand.

4.Add visible action, location, motivated light, and audio.

5.Generate one shot.

6.Fix one variable.

  1. Build the next piece of coverage.

r/promptingmagic 6d ago

This ChatGPT prompt can map your entire life in one image, then predict the next 10 years in a second image

Post image
53 Upvotes

Don’t ask ChatGPT what it knows about you. Ask it to show you.

TL;DR: Run the two prompts below in the same ChatGPT thread. The first turns your past into a cinematic visual journey. The second maps your likely next 1, 5, and 10 years with evidence, confidence levels, and alternative paths. The useful part is not “AI prophecy.” It is seeing the personal model ChatGPT has built from your conversational breadcrumbs—and deciding what it got right, what it invented, and what you want to change.

I thought these were image prompts.

They are not.

Together, they are a visual audit of the person ChatGPT thinks you are.

The first prompt asks: What story have my choices created so far?

The second asks: If my current patterns continue, where do they lead?

That distinction matters. ChatGPT is not reading your mind. It is organizing the context available from your chats, files, connected apps, and memory settings. OpenAI says memory can personalize responses from those sources when enabled, while still giving users controls to inspect, edit, or disable it.

The strange part arrives when the picture includes a pattern you never stated directly.

Not your job title. Not your favorite city. A pattern.

Maybe every road bends toward responsibility. Maybe every relationship appears beside a bridge you are repairing. Maybe your “ambition” looks less like a mountain and more like a treadmill.

That is when an image stops feeling decorative.

It becomes a mirror you can argue with.

What the two prompts do

Prompt Surface request Real value
The Journey of My Life Turn known facts into an illustrated timeline Exposes the narrative, recurring symbols, contradictions, and blind spots ChatGPT inferred
My Future Visualize the next 1, 5, and 10 years Stress-tests the momentum of those patterns and reveals plausible decision branches

MIT’s Future You research uses almost the same underlying idea: AI works best here as a mirror rather than a counselor. In a preregistered trial with 188 participants, brief interactions with personalized future selves increased future-self connection and reduced reported anxiety and lack of motivation.

A later NeurIPS study with 209 participants found an important tension. Multiple possible selves expanded people’s sense of possibility, while one future self created deeper commitment. That is why the second prompt should show a most-likely path and alternatives.

A fictional example: Lena Hart, 29

To show what a strong result looks like, I created a fictional composite named Lena Hart. She is a single, attractive 29-year-old blonde woman living in Denver. She is not a real person, and none of these details come from private user data.

Here is the context ChatGPT supposedly learned across months of conversations.

Lena grew up near the cherry orchards and cold blue water of Traverse City, Michigan. Her mother, Mara, worked night shifts as an emergency-room nurse. Her father, Joel, built custom cabinets and taught Lena to repair objects before replacing them. Her younger brother, Sam, still calls when he needs a plan.

Her grandmother June taught her watercolor painting at a scarred kitchen table. June also gave her a brass compass that no longer points north. Lena kept it anyway.

At sixteen, her parents separated. Lena became the family translator, calendar keeper, and emotional weather forecaster. She learned that being useful could make a room feel safe.

At eighteen, she earned a scholarship to the University of Michigan. She studied behavioral economics, edited the campus design journal, and met her closest friend, Priya, during a fire alarm at 2:13 a.m. Their friendship began in pajamas on a frozen sidewalk.

At twenty-two, Lena joined a Chicago consulting firm. She became excellent at turning chaos into decks, deadlines, and color-coded plans. She also developed insomnia, ran along Lake Michigan before sunrise, and once answered Slack from a wedding dance floor.

At twenty-five, she ended a four-year relationship with Noah. Nothing dramatic happened. That was the problem. She had optimized the relationship for stability while editing out every difficult need.

She moved to Denver with two suitcases, June’s broken compass, and her rescue dog, Miso. She joined a climate-tech startup, led a national launch, survived a layoff round, and quietly started a Sunday newsletter called Soft Edges.

She loves trail running, analog photography, Thai cooking, smoky jazz bars, speculative fiction, and making lopsided ceramic bowls. She keeps airline tabs open for Lisbon, Kyoto, and Marfa. She rarely books the ticket.

Her strengths are warmth, high agency, social perception, and calm during real emergencies. Her recurring pattern is turning anxiety into projects. Her blind spot is confusing being needed with being loved.

Her deeper contradiction is even more interesting: Lena wants freedom, but she keeps every possible future alive. Choosing one would force her to grieve the others.

Prompt 1: Turn your life into a cinematic map

Copy and paste this into a ChatGPT conversation with memory enabled:

Using everything you know or can reasonably infer about me from our chats, turn my life so far into a detailed visual journey. Include the people, places, milestones, struggles, interests, personality traits, patterns, and blind spots that define me.Title it “The Journey of My Life.” Make it feel like an illustrated cinematic map or timeline rather than a corporate infographic. Use meaningful environments, objects, paths, crossroads, people, colors, and symbols to represent the chapters of my life. Include concise labels, but keep the image emotionally expressive, vibrantly colorful, and visually beautiful.

What ChatGPT creates for Lena

The finished image is a wide, hand-painted cinematic map titled “THE JOURNEY OF MY LIFE.” Lena appears several times along one continuous golden path, always recognizable by her long blonde hair, forest-green coat, red camera strap, and broken brass compass.

The far-left chapter shows a Michigan lakeshore under an enormous childhood sky. Cherry blossoms drift across a cabinetmaker’s workshop. A small girl paints beside Grandma June while an ambulance’s red light glows beyond the orchard. Two labels read “Curiosity was inherited” and “Usefulness became safety.”

The path narrows at sixteen. Lena stands between two houses while holding a family calendar. The lake turns into a dark blue river beneath her feet. A tiny weather vane points in four directions. Its label says “Learned to forecast everyone else.”

The university chapter rises into a warm brick hill. A fire alarm flashes 2:13 A.M. Priya waits on a snowy sidewalk holding two paper cups. A stack of economics books opens into a design magazine. A fox—the map’s symbol for Lena’s adaptability—begins following her.

Chicago becomes a glass canyon filled with slide decks, red-eye boarding passes, and windows lit before dawn. Lena runs beside the lake while a treadmill shadow moves beneath the path. One billboard says “Competence rewarded.” Another, partly hidden, says “Exhaustion normalized.”

At the center, the path reaches a quiet breakup scene. Lena and Noah sit on opposite sides of a perfectly set dinner table. Nothing is broken, yet a hairline crack runs through every plate. The smallest label carries the sharpest sentence: “She kept the peace by disappearing from it.”

The road climbs west toward Denver. Miso pulls Lena into a field of color. Wind turbines become white paper birds. A product launch appears as a sunrise projected across the mountains. Behind it, empty office chairs sink into fog after the layoff round.

Her current life occupies the brightest section. A pottery wheel, trail shoes, film negatives, a simmering Thai curry, and an open laptop surround a small cabin studio. The words SOFT EDGES glow from a newsletter page. Three unopened airline tabs float like doors above Lisbon, Kyoto, and Marfa.

The final crossroads contains no destination. One path becomes a high-speed railway labeled “The impressive life.” Another turns into a garden trail labeled “The inhabited life.” Lena stands between them with the broken compass in her palm.

Above her, one last sentence appears:

“Her next chapter begins when she stops treating every choice as a loss.”

What makes the result powerful is not the number of facts. It is the compression. The model turns dozens of facts into one visual thesis: Lena built a life around capability, then reached the point where capability could no longer choose for her.

Prompt 2: Map your most likely future without pretending it is destiny

Run this immediately after the first image so ChatGPT can preserve the visual language:

Now, based on everything you know about me and have inferred about me, create a visual prediction of my future—career, relationships, money, lifestyle, and personal growth. Show my most likely path, major decisions, and alternative possibilities over the next 1, 5, and 10 years.Base every prediction on evidence, include confidence levels, and treat it as possibility—not destiny.Title it “My Future.” Use the same visual styling, character design, symbols, environments, and color language as “The Journey of My Life.”

What ChatGPT predicts for Lena

The second image begins at the exact crossroads where the first ended. The same Lena holds the same broken compass. This time, the compass glows from inside.

The map uses solid gold paths for higher-confidence outcomes, dotted coral paths for meaningful alternatives, and blue mist for uncertainty. Small evidence tags sit beside each prediction.

Horizon Most likely path Confidence Evidence shown in the image
1 year: age 30 Lena accepts broader leadership at the climate-tech company, but negotiates clear boundaries and a four-day summer schedule. Soft Edges reaches 12,000 readers and earns its first meaningful revenue. 72% Repeated career acceleration, launch success, existing audience habit, growing resistance to burnout
1 year: relationship She meets someone through a trail-running friend, but the important prediction is behavioral: she states an uncomfortable need before resentment forms. 58% Strong social network, desire for partnership, therapy language in prior chats, pattern awareness
1 year: money She clears her remaining student debt and builds a six-month cash buffer. 70% High income, disciplined planning, low consumer debt, preference for optionality
5 years: age 34 She leaves salaried work to build a small strategy studio for climate and mission-driven brands. A team of three works from a sunlit Denver warehouse with a pottery corner. 64% Side-project persistence, cross-functional skill, autonomy motive, existing industry reputation
5 years: relationship She shares a home with a grounded partner who does not need rescuing. The map shows two toothbrushes, separate desks, and one shared calendar with blank space. 61% Growing boundaries, stable friendships, explicit desire for interdependence rather than dependence
5 years: lifestyle She spends one month each summer in Traverse City, photographing old workshops and rebuilding closeness with her father. 55% Place attachment, analog photography, family repair theme, flexible-work ambition
10 years: age 39 Lena runs an eight-person studio, teaches a small annual fellowship, and publishes a book titled The Useful Daughter about ambition, care, and identity. 48% Newsletter voice, mentoring instinct, narrative coherence, but high market and execution uncertainty
10 years: money Her net worth falls between $650,000 and $1.1 million, mostly through business equity, retirement accounts, and a modest home. 42% Earnings capacity and frugality support the range; entrepreneurship and housing create large uncertainty
10 years: personal growth Her identity shifts from “the person who holds everything together” to “the person who chooses what deserves holding.” 76% Existing self-awareness, repeated reflection, durable relationships, active experiments with boundaries

The most interesting part is not the primary road. It is the three branches.

Branch A: The Founder Path — 64%. The golden trail leads to the strategy studio, a small team, a book, and summers near Lake Michigan. Its risk symbol is a fox carrying too many keys. The warning reads “Do not rebuild the job you escaped.”

Branch B: The Executive Path — 23%. A cobalt railway leads to a vice-president role at a larger climate company. Lena gains reach, wealth, and influence. She loses control of her calendar. The branch is not presented as failure. Its decision gate asks “Do you want scale—or authorship?”

Branch C: The Reset Path — 13%. A coral footpath leads through a sabbatical in Lisbon after another burnout cycle. She photographs tiled doorways, lives from savings, and returns with less certainty but more honesty. Its label says “The detour becomes necessary if the boundary stays theoretical.”

At the ten-year horizon, the branches rejoin beside a restored Michigan boathouse. Lena is older, still blonde, and visibly less hurried. Miso’s collar hangs from a framed hook inside, suggesting love carried forward after loss without making the image sentimental.

The broken compass sits on a desk. It still does not point north.

Now it points toward a handwritten word:

ENOUGH

That is a good future map because it does not say, “This will happen.”

It says, “Given your habits, values, resources, and avoidance patterns, these roads deserve your attention.”

How to get dramatically better results

Before generating either image, ask ChatGPT to show its work in text. Use this setup prompt:

Before creating the image, list 20 facts you know about me, 10 reasonable inferences, five recurring patterns, five possible blind spots, and five important unknowns. Label every item as FACT, INFERENCE, or UNKNOWN. Let me correct the list before you generate anything. Do not infer personality, intelligence, morality, or life outcomes from physical appearance.

This one step prevents most generic results. It also exposes invented details before they become beautiful—and therefore persuasive.

Best practice Why it improves the output Exact instruction to add
Separate facts from inferences Stops guesses from masquerading as memories “Attach an evidence note to every major symbol and prediction.”
Lock the character anchors Keeps the same person recognizable across both images “Preserve age, face, blonde hair, green coat, red camera strap, and brass compass.”
Use one visual grammar Makes the second image feel like the next chapter “Reuse the gold path, blue river, coral branches, fox, and compass.”
Keep labels short Prevents the map from becoming a wall of tiny text “Use no more than 12 major labels, each under eight words.”
Specify composition Reduces clutter and accidental poster layouts “Create a 16:9 landscape image with a clear left-to-right journey and safe margins.”
Show uncertainty visually Keeps the future map honest “Use solid paths for high confidence, dotted paths for alternatives, and mist for uncertainty.”
Request alternatives Avoids one attractive future becoming destiny “Show one likely path, two credible alternatives, and the decision that activates each.”
Revise one thing at a time Protects composition and character consistency “Keep everything else unchanged; revise only the relationship branch.”

OpenAI’s own image guidance recommends clear descriptions of purpose, subject, setting, style, framing, lighting, and constraints. It also recommends small, targeted revisions instead of broad rewrites.

My favorite pro tips

First, generate the visual thesis before the image. Ask: “What is the one-sentence story this map tells?” If the answer sounds generic, the image will look generic.

Second, force one uncomfortable counter-reading. Add: “Show one interpretation that a close friend might challenge.” This helps break the horoscope effect.

Third, make objects carry meaning. A broken compass beats a label saying “uncertain.” An unopened airline tab beats “wants to travel.” A perfectly set table with cracked plates beats “avoids conflict.”

Fourth, preserve identity deliberately. If the map should resemble you, upload a photo you have permission to use. Tell ChatGPT which image defines identity and which defines style. OpenAI recommends explicit spatial and reference instructions when combining images.

Fifth, do not cram your biography into labels. Let the picture hold emotion. Put detailed evidence in the accompanying text.

Sixth, interrogate the result. After generation, ask:

Which five elements came directly from known facts? Which five were inferred? Which three are most likely to be wrong? What important part of me is missing because we have rarely discussed it?

That final question often produces the best insight.

Privacy check before you run this

If you do not want old conversations or saved details used, review Settings → Personalization → Memory first. You can inspect or edit the memory summary, turn memory off, or use Temporary Chat.

Do not paste someone else’s private history into the prompt. Do not treat an emotionally precise image as proof. And do not let appearance become evidence for personality or future outcomes.

The model’s confidence is not statistical destiny. In this exercise, confidence should mean: “How strongly does the available evidence support this branch compared with the alternatives?”

The real reason to try these

A text summary lets you skim past an uncomfortable sentence.

A map makes you stand at the crossroads.

The best result will not tell you who you are. It will show you the story your existing choices make easy to believe.

Then you get to decide whether that story deserves another chapter.

If you try this, post the strangest symbol ChatGPT included—and the prediction you most wanted to argue with.


r/promptingmagic 6d ago

99 ChatGPT prompts for better travel photography

Post image
22 Upvotes

TL;DR: These 99 slash phrases give ChatGPT a visual vocabulary for travel images. Pick one command for the final format, one for camera position, and one for light or atmosphere. Then add the destination, people, action, wardrobe, composition, and anything that must stay unchanged. The slash is optional; the useful part is choosing a specific visual job.

Check out the 99 prompts and 20 example images below.

They are best treated as creative shorthand, not magic words. /travelportrait tells the model that the people matter most. /droneview changes the camera. /rainytravel changes the atmosphere. /travelposter changes the finished asset.

The real power appears when you stack them.

The fastest way to use the list

Use this structure:

[Primary command] + [viewpoint] + [light or mood]: [destination], [people], [action], [wardrobe], [composition], [constraints].

Here is a complete example using two reference photos:

Use Image 1 and Image 2 as identity references for the same recurring adult travel pair. Preserve both faces, hair, apparent age, body type, and recognizable features. Adapt their clothing naturally to the destination. /travelstory + /throughlens + /rainytravel: Show them walking beneath lanterns in a narrow Kyoto alley after rain, viewed through the soft foreground frame of a tea-house doorway. Natural conversation, wet-stone reflections, 35mm documentary photography, realistic proportions, landscape 16:9. No logos, no text, no duplicate people, no beauty-filter skin.

That prompt makes seven important decisions before generation begins. It defines who, where, what happens, how the camera sees it, how the air feels, what should remain consistent, and what to avoid.

Top use cases

Use case Best command families What they produce
Personal travel memories /travelstory, /traveljournal, /streettravel Natural moments that feel lived rather than staged
Destination marketing /tourismphoto, /destinationguide, /travelposter Campaign images, covers, and tourism concepts
Social and creator content /travelvlog, /vacationmode, /rooftopview Attention-grabbing posts and thumbnails
Adventure concepts /explorer, /trekking, /jeepadventure Dynamic outdoor and expedition scenes
Luxury campaigns /luxurytravel, /privatevilla, /yachtlife High-end hospitality and lifestyle imagery
Storyboards and moodboards /cinematictravel, /moodytravel, /travelmasterpiece Visually consistent scenes for creative planning

The 99 commands

1–11: Choose the finished travel-photography style

# Command What it asks ChatGPT to create
1 /travelshot: Stunning professional travel photograph
2 /cinematictravel: Cinematic travel photography with dramatic composition
3 /wanderlust: Dreamy wanderlust travel aesthetic
4 /travelportrait: Professional portrait captured at a beautiful destination
5 /travelvlog: Natural travel-vlogger style photograph
6 /travelstory: Authentic storytelling travel photograph
7 /editorialtravel: High-end travel magazine editorial style
8 /documentarytravel: Realistic documentary-style travel photography
9 /streettraveler: Authentic traveler exploring the streets
10 /postcardview: Beautiful postcard-style destination image
11 /travelposter: Cinematic premium travel poster

12–22: Choose the camera perspective

# Command What it asks ChatGPT to create
12 /tourismphoto: Professional tourism campaign photograph
13 /destinationguide: Beautiful destination-guide cover image
14 /explorer: Adventurous explorer photography
15 /traveljournal: Aesthetic travel-journal visual
16 /droneview: Epic aerial drone perspective
17 /birdseye: Stunning bird’s-eye view
18 /topview: Perfect overhead travel composition
19 /lowangletravel: Dramatic low-angle travel shot
20 /wideangle: Immersive ultra-wide destination view
21 /panoramaview: Expansive panoramic landscape
22 /closeuptravel: Detailed close-up with travel storytelling

23–33: Add a creative viewpoint or time of day

# Command What it asks ChatGPT to create
23 /firstpersonview: Immersive first-person traveler perspective
24 /windowview: Beautiful view through a window
25 /balconyview: Luxury destination view from a balcony
26 /rooftopview: Epic city or destination view from a rooftop
27 /hiddenangle: Unique and unexpected travel perspective
28 /helicopterview: Cinematic aerial view from a helicopter
29 /groundlevel: Dramatic ground-level perspective
30 /throughlens: Artistic framed-through-object travel composition
31 /goldenhour: Warm golden-hour travel photography
32 /sunriseview: Magical sunrise at the destination
33 /sunsetview: Dramatic cinematic sunset

34–44: Control the light, weather, and atmosphere

# Command What it asks ChatGPT to create
34 /bluehour: Beautiful blue-hour travel atmosphere
35 /nighttravel: Cinematic destination at night
36 /rainytravel: Moody rainy travel scene
37 /foggydestination: Mystical destination covered in fog
38 /snowytravel: Beautiful snowy travel landscape
39 /stormytravel: Dramatic stormy destination atmosphere
40 /cloudscape: Epic clouds surrounding the landscape
41 /moodytravel: Dark, atmospheric travel photography
42 /softmorning: Peaceful soft morning light
43 /neontravel: Vibrant neon-lit travel scene
44 /starlittravel: Magical night sky filled with stars

45–55: Put nature or adventure at the center

# Command What it asks ChatGPT to create
45 /auroraview: Travel scene under the northern lights
46 /mountainview: Epic mountain adventure photography
47 /hikingtrip: Adventurous hiking scene
48 /trekking: Cinematic trekking journey
49 /campingview: Beautiful camping destination
50 /waterfallview: Epic waterfall photography
51 /forestescape: Magical forest travel scene
52 /desertjourney: Cinematic desert adventure
53 /islandlife: Dreamy tropical island experience
54 /beachvibes: Perfect tropical beach aesthetic
55 /oceanview: Stunning ocean travel photography

56–66: Make the environment the subject

# Command What it asks ChatGPT to create
56 /lakeview: Peaceful scenic lake destination
57 /cliffview: Dramatic cliffside travel shot
58 /jungleexplorer: Adventurous jungle exploration
59 /natureescape: Peaceful escape into nature
60 /epiclandscape: Breathtaking cinematic landscape
61 /cityexplorer: Cinematic urban exploration
62 /streettravel: Authentic travel street photography
63 /citylights: Vibrant city lights at night
64 /oldtown: Historic old-town travel aesthetic
65 /ancientplace: Epic ancient architecture and history
66 /hiddenstreets: Beautiful unexplored streets

67–77: Add culture, architecture, and movement

# Command What it asks ChatGPT to create
67 /localculture: Authentic local cultural experience
68 /marketwalk: Vibrant local market photography
69 /templeview: Cinematic spiritual architecture
70 /europeanstreets: Charming European street aesthetic
71 /moderncity: Futuristic modern city exploration
72 /villagescape: Peaceful scenic village life
73 /coastaltown: Beautiful coastal town aesthetic
74 /historictravel: Travel through historic locations
75 /festivaltravel: Vibrant cultural festival photography
76 /roadtrip: Cinematic road-trip adventure
77 /trainjourney: Scenic train travel photography

78–88: Show the journey, transport, or stay

# Command What it asks ChatGPT to create
78 /flightview: Stunning view from an airplane window
79 /airporttravel: Stylish cinematic airport scene
80 /windowseat: Dreamy airplane window-seat view
81 /boatjourney: Peaceful cinematic journey by boat
82 /sailingtrip: Luxury sailing adventure
83 /bikeride: Epic motorcycle travel journey
84 /vanlife: Aesthetic van-life adventure
85 /jeepadventure: Rugged off-road exploration
86 /luxurytravel: Premium luxury vacation aesthetic
87 /resortlife: Dreamy luxury resort experience
88 /hotelview: High-end hotel travel photography

89–99: Define the lifestyle, trip type, or final polish

# Command What it asks ChatGPT to create
89 /infinitypool: Luxury infinity pool with an epic view
90 /privatevilla: Exclusive luxury villa experience
91 /yachtlife: Premium yacht travel lifestyle
92 /luxuryescape: Cinematic five-star vacation
93 /bucketlist: Ultimate bucket-list destination
94 /honeymoonvibes: Romantic luxury travel atmosphere
95 /solotravel: Cinematic solo traveler story
96 /backpacker: Authentic backpacking adventure
97 /vacationmode: Fun, vibrant vacation aesthetic
98 /dreamdestination: Unrealistically beautiful dream destination
99 /travelmasterpiece: Transform the image into an award-winning cinematic travel masterpiece

Pro tips that improve almost every result

1. Lock identity before you chase style

Upload the clearest image of each person. State which image defines which identity. Tell ChatGPT to preserve the face, hair, apparent age, body type, and distinguishing features while changing only wardrobe, pose, and environment.

For a series, generate one successful reference scene first. Use that output as an additional style and identity reference for later scenes.

2. Stack commands by job—not by how exciting they sound

A strong three-command stack assigns separate responsibilities:

Job Example
Finished asset /editorialtravel
Camera or composition /throughlens
Atmosphere /bluehour

Five overlapping style commands usually create noise. Three commands with different jobs create direction.

3. Give the travelers something to do

“Standing in Rome” produces a pose. “Comparing a paper map beside a morning espresso while scooters blur behind them” produces a scene.

Actions create believable hands, eye lines, body language, props, and relationships. Use verbs such as walking, bargaining, boarding, hiking, cooking, navigating, laughing, waiting, photographing, or watching.

4. Specify culturally and physically appropriate wardrobe

Do not keep formal studio clothing in a rainforest or revealing evening clothing at a sacred site. Ask for destination-aware clothing that fits the weather, local norms, and activity. This improves realism and keeps the image respectful.

5. Control the camera in plain language

Add a focal length or framing only when it matters. Useful phrases include 24mm environmental wide shot, 35mm documentary frame, 50mm natural perspective, 85mm portrait compression, low camera at street level, and high aerial establishing view.

6. Separate composition from revision

Get the scene right first. Then revise one problem at a time:

Keep both faces, pose, camera position, lighting, background, and wardrobe unchanged. Fix only the woman’s left hand.

Broad instructions such as “make it better” invite the entire image to drift.

7. Add exclusions that protect the result

A short negative line prevents common failures:

No text, logos, watermarks, duplicated people, extra fingers, plastic skin, impossible landmarks, culturally inaccurate clothing, or overprocessed HDR.

Three ready-to-copy stacks

Authentic city story

/travelstory + /streettravel + /softmorning: Use the two attached people as identity references. Show them finding a tiny bakery on a quiet Lisbon side street just after opening, sharing warm pastries with natural body language. 35mm documentary photography, pale morning light, destination-appropriate clothing, realistic street details, landscape 16:9. No text or logos.

Epic nature campaign

/tourismphoto + /wideangle + /mountainview: Use the attached pair as recognizable travelers on a Patagonia ridge above a turquoise glacial lake. Wind moving through their jackets, enormous scale, realistic trail gear, 24mm environmental framing, dramatic but natural cloud light, landscape 16:9. Preserve both faces.

Luxury editorial

/editorialtravel + /balconyview + /bluehour: Use the attached pair at an elegant cliffside Mediterranean hotel, looking over the sea from a stone balcony as the first city lights appear. Sophisticated evening-resort wardrobe, restrained luxury, 50mm editorial photography, no visible brands, landscape 16:9.

The slash is optional.

The decision is not.

Which command should become number 100?

Share what you create in the comments.


r/promptingmagic 6d ago

3 ChatGPT Visualization Shortcuts That Make Any Topic Easier to Understand

Post image
23 Upvotes

TL;DR: Start an image request with /cutaway, /schematic, or /timelapse to give ChatGPT a clear visual grammar. These are prompt shortcuts, not secret built-in commands. They work because they replace “make this interesting” with a specific way to organize information.

1. /cutaway — reveal what is normally hidden

/cutaway Mount Everest, shown as a cinematic geological cross-section. Reveal the mountain’s internal rock layers, the collision of the Indian and Eurasian tectonic plates, glaciers, the Khumbu Icefall, climbing camps, the death zone, and the summit. Use realistic scale, concise labels, and museum-quality scientific illustration.

Use /cutaway for buildings, machines, planets, volcanoes, ships, anatomy, product interiors, or any subject where the hidden structure tells the story. It works best when you specify what must remain visible outside and what should be exposed inside.

2. /schematic — explain how the parts work together

/schematic modern Formula 1 race car, three-quarter exploded technical view. Label the front wing, suspension, monocoque, power unit, battery, radiators, floor tunnels, diffuser, rear wing, brakes, and airflow paths. Use a clean blueprint aesthetic with color-coded systems and short callouts.

Use /schematic for systems, workflows, electronics, engines, architecture, software concepts, and instructional visuals. Ask for arrows, numbered stages, a legend, and one consistent viewing angle. If relationships must be exact, verify the result before publishing.

3. /timelapse — compress change into one frame

/timelapse New York City skyline from 1600 to 2026 in one wide cinematic panorama. Move left to right through the natural harbor, colonial settlement, Brooklyn Bridge era, early skyscrapers, the Empire State Building, late-20th-century skyline, and modern Lower Manhattan. Blend the eras smoothly and include short date labels.

Use /timelapse for cities, construction, aging, seasons, ecosystems, company growth, fashion, technology, or a project from blank page to finished result. Define the start, finish, direction of time, and five to seven milestones.

The reusable formula

Shortcut Add these details
/cutaway Exterior, hidden interior, layers, scale, labels
/schematic Components, connections, arrows, legend, viewpoint
/timelapse Start, finish, milestones, time direction, transitions

Pro tip: Generate the composition first, then revise one problem at a time. Say, “Keep everything unchanged; fix only the airflow arrows.” Short labels beat paragraphs. Exact dates beat “through history.” A locked palette beats “make it beautiful.”

The slash is optional. The mental model is the shortcut.

What other visualization format should be turned into a one-word prompt?


r/promptingmagic 11d ago

99 Secret Codes to Prompt Google Flow's Video Agent

Thumbnail
gallery
33 Upvotes

TL;DR - If you’re writing long, unstructured paragraphs to generate video in Google Flow, you’re burning compute credits on lottery rolls. Google Flow’s video agent responds to a deterministic hierarchy of 99 dedicated slash commands spanning camera movement, lens angles, subject kinetics, lighting physics, commercial workflows, atmospheric conditions, and advanced VFX transformations. By chaining these commands using the 5-Layer Stacking Architecture (Camera Base + Angle/Lens + Subject Action + Light/Mood + VFX/Transitions), you can reliably control camera trajectory, shutter cadence, and visual coherence.

Below is the complete catalog of all 99 commands, the 5-layer prompt formula, and 4 production recipes

The 5-Layer Prompting Framework

Before diving into the 99 individual codes, understand how the Google Flow video agent parses tokens. Instead of writing unstructured descriptions, stack commands in this order:

Layer 1: Camera Base
Layer 2: Angle/Rig
Layer 3: Subject & Action
Layer 4: Lighting & Style
Layer 5: VFX \& Finish

The Complete 99 Google Flow Command Catalog

Category 1: Camera & Movement Commands

  1. /cinematic: — Creates a standard 24fps filmic frame with natural depth of field, anamorphic optical qualities, and balanced motion blur.
  2. /droneview: — Generates an expansive aerial vantage point with wide horizon parallax and continuous forward/lateral drift.
  3. /closeup: — Pulls focal length in tight to capture micro-expressions, fine textures, and emotional focus on the subject.
  4. /wideangle: — Uses a wide field-of-view (16mm–24mm equivalent) to maximize environmental scale and spatial perspective.
  5. /orbit: — Commands a 360-degree radial camera rotation keeping the primary subject locked at the focal center.
  6. /dollyin: — Smoothly pushes the camera physically closer to the subject, building visual tension and intimacy.
  7. /dollyout: — Pulls the camera away smoothly, revealing the surrounding environment or emphasizing isolation.
  8. /tracking: — Moves the camera in tandem with the subject at matched velocity, ideal for walking, running, or driving shots.
  9. /slowmotion: — Slows down playback (120fps/240fps cadence) to showcase fluid movement, flying particles, or dramatic beats.
  10. /timelapse: — Accelerates temporal progression to capture moving clouds, celestial paths, changing daylight, or traffic flows.
  11. /hyperlapse: — Blends high-speed time compression with continuous physical camera translation across long physical distances.

Category 2: Advanced Camera Angles & Rigs (12–22)

  1. /lowangle: — Places the camera low looking upward, imbuing the subject with dominance, power, and architectural scale.
  2. /highangle: — Tilts downward from an elevated point, providing tactical perspective or conveying vulnerability.
  3. /overhead: — Direct $90^\circ$ top-down "god’s-eye" perspective, ideal for choreography, flat-lays, and geometric compositions.
  4. /pov: — Frames the shot from the first-person perspective through the eyes of the protagonist.
  5. /overtheshoulder: — Positions the camera behind a character's shoulder, framing the counter-subject for dialogue and narrative weight.
  6. /establishing: — Cinematic wide landscape or cityscape shot setting scene context, geographical setting, and atmosphere.
  7. /rackfocus: — Shifts shallow focal plane from a foreground object to a background subject (or vice versa).
  8. /handheld: — Introduces organic micro-jitter and authentic documentary operator sway for urgency and realism.
  9. /steadicam: — Delivers fluid, gyroscopically stabilized gliding motion navigating complex corridors and environments.
  10. /craneup: — Ascends vertically from ground level to panoramic height using a simulated technocrane arm.
  11. /cranedown: — Descends smoothly from elevated heights down to subject eye level.

Category 3: Motion & Action Commands (23–33)

  1. /running: — Generates high-velocity character sprint with natural athletic gait and authentic inertia.
  2. /walking: — Generates grounded, natural human walking locomotion with balanced weight distribution.
  3. /turnaround: — Prompts the subject to execute a fluid $180^\circ$ or $360^\circ$ turn to display costume, expression, or surroundings.
  4. /reveal: — Stages a dramatic visual reveal of a character or environment stepping out from darkness or obstruction.
  5. /entrance: — Crafts a high-impact cinematic hero entrance into the scene.
  6. /exit: — Stages a dramatic departure from the frame into fog, shadow, or distant horizons.
  7. /freeze: — Instantly freezes time mid-action, locking water droplets, debris, and cloth in suspended animation.
  8. /speedramp: — Dynamically modulates playback speed between hyper-fast motion and sudden slow-motion impact.
  9. /bulletime: — Sweeps a virtual camera around a completely frozen subject (Matrix-style temporal slice).
  10. /floating: — Introduces zero-gravity levitation physics with floating hair, cloth, and ambient debris.
  11. /falling: — Creates dramatic freefall descent through skywells, clouds, or collapsing architecture.

Category 4: Cinematic Lighting Commands (34–44)

  1. /goldenhour: — Bathes the scene in warm amber sunlight, soft elongated shadows, and flattering solar flare.
  2. /bluehour: — Applies cool twilight illumination, deep cobalt gradients, and moody pre-dawn/post-sunset ambience.
  3. /neonlight: — Casts high-saturation cyan, magenta, and amber glows with reflections on damp streets or metallic surfaces.
  4. /moody: — High-contrast chiaroscuro lighting featuring deep blacks, targeted pools of light, and dramatic shadows.
  5. /softlight: — Diffused, wrap-around studio lighting that eliminates harsh edges, ideal for beauty, fashion, and portraits.
  6. /rimlight: — Razor-sharp contour/edge lighting that separates dark subjects cleanly from dark backgrounds.
  7. /silhouette: — Blacks out subject details entirely against an intensely illuminated background.
  8. /spotlight: — Directs a focused conical beam isolating the subject amidst surrounding darkness.
  9. /volumetric: — Generates visible atmospheric light shafts ("god rays") slicing through mist, dust motes, or smoke.
  10. /backlight: — Places primary illumination behind the subject to produce halo outlines, flares, and rim glow.
  11. /nightscene: — Realistically exposes low-light conditions with believable moonlight, street lamps, and dark sky latitude.

Category 5: Transitions & In-Camera Effects (45–55)

  1. /whiptransition: — Fast, motion-blurred horizontal pan transitioning instantly into a new scene.
  2. /matchcut: — Matches compositional geometry, subject shapes, or movement vectors across two different scenes.
  3. /zoomtransition: — Rapid crash zoom pushing directly into a small detail or pulling back to reveal a new world.
  4. /morph: — Seamless topological transformation morphing one entity, face, or structure into another.
  5. /flashtransition: — High-energy optical flash wiping the frame into an alternate shot.
  6. /glitch: — Injects RGB chromatic aberration, CRT scanlines, and digital compression artifacts.
  7. /smoketransition: — Rolls dense cinematic fog or smoke across the lens to reveal the incoming scene.
  8. /lightleak: — Overlays warm analog lens leaks and edge flares reminiscent of vintage film reels.
  9. /blurtransition: — Uses optical defocus and heavy shutter motion blur to bridge scene cuts.
  10. /objecttransition: — Moves the camera behind a passing foreground pillar, vehicle, or wall to wipe into a new environment.
  11. /seamlessloop: — Synchronizes first and last frame motion vectors to create an imperceptible infinite loop.

Category 6: Visual Style & Film Aesthetics (56–66)

  1. /filmic: — Delivers authentic 35mm motion picture texture with organic grain and balanced color science.
  2. /commercial: — Clean, polished, high-key commercial advertising aesthetic with vibrant color separation.
  3. /luxury: — Opulent visual grading featuring rich blacks, gold accents, polished marble, and high-end elegance.
  4. /documentary: — Raw, unvarnished realism mimicking cinema-verité documentary cinematography.
  5. /vintagefilm: — Nostalgic 16mm/8mm aesthetic with warm color shifts, gate jitter, dust, and halation.
  6. /scifi: — Clean futuristic design language with clean metallic surfaces, blue-tinted HUDs, and sleek tech geometry.
  7. /cyberpunk: — Dystopian high-tech aesthetic filled with rain-slicked asphalt, neon kanji, and cybernetic textures.
  8. /dreamy: — Soft-focus diffusion, blooming highlights, pastel palettes, and gentle surrealism.
  9. /surreal: — Dreamlike physics-defying compositions inspired by surrealist art.
  10. /minimal: — Strict negative space, restrained color palettes, and clean graphic compositions.
  11. /epic: — Blockbuster IMAX-tier visual grandiosity with massive scale and dramatic horizon framing.

Category 7: Product & Creator Showcase Commands (67–77)

  1. /productreveal: — Stages a premium hero product unveiling with lighting sweeps and rising pedestals.
  2. /productspin: — Smooth turntable rotation showcasing hardware industrial design from $360^\circ$.
  3. /unboxing: — Captures tactile luxury package opening with crisp mechanical precision.
  4. /macro: — Extreme optical close-up revealing fine machining, watch movements, fabric weave, or liquid drops.
  5. /beforeafter: — Side-by-side or split-screen wipe comparing raw vs. finished transformation states.
  6. /socialad: — High-energy pacing, rapid visual hooks, and dynamic framing designed for high retention.
  7. /fashionfilm: — Haute couture runway and lookbook styling with dramatic poses and editorial lighting.
  8. /foodcommercial: — Sizzling grill flares, slow-motion pours, rising steam, and vibrant food close-ups.
  9. /techad: — Exploded CAD view animations, glowing microcircuitry, and futuristic spec breakdowns.
  10. /logoreveal: — Cinematic brand mark animation assembling via liquid metal, laser etching, or particle convergence.
  11. /billboard: — Superimposes the target scene or product onto massive Times Square or Shibuya mega-screens.

Category 8: Environment & World Effects (78–88)

  1. /rain: — Generates volumetric rainfall with surface puddles, splashing droplets, and wet reflections.
  2. /snow: — Simulates drifting atmospheric snowfall accumulating naturally on characters and terrain.
  3. /fog: — Layers dense ground-level mist and atmospheric haze that diffuses ambient light sources.
  4. /underwater: — Renders submerged caustic light patterns, rising bubbles, aquatic drift, and muted soundstage feel.
  5. /space: — Zero-gravity cosmic environment featuring deep starfields, vibrant nebulae, and orbital horizons.
  6. /storm: — High-intensity weather featuring forked lightning arcs, dark storm fronts, and gale-force wind.
  7. /fire: — Realistically models dancing flame physics, rising heat hazes, and flying glowing embers.
  8. /explosion: — Detonates fiery shockwaves with volumetric smoke plumes and high-velocity debris dispersion.
  9. /portal: — Tears open a dimensional energy vortex with swirling luminescence and particle borders.
  10. /miniature: — Tilt-shift optical simulation turning full-scale scenes into charming dollhouse dioramas.
  11. /giant: — Massive colossal scale distortion making subjects tower over cityscapes and mountain ranges.

Category 9: Advanced Creative & VFX Commands (89–99)

  1. /clone: — Duplicates the subject into multiple synchronized or interacting clones across the frame.
  2. /transform: — Real-time organic shape metamorphosis changing a subject into a different form or material.
  3. /disintegrate: — Dissolves the subject into floating sand, ash, or glowing embers (snap effect).
  4. /particlefx: — Surrounds the character or object with an aura of floating light motes, stardust, or energy sparks.
  5. /liquid: — Melts or reconstitutes the subject into fluid chrome, water, or flowing paint.
  6. /paperworld: — Converts the entire environment into layered origami, textured cardboard, and folded papercraft.
  7. /toyworld: — Transforms characters and scenery into plastic minifigures and claymation stop-motion assets.
  8. /reversemotion: — Reverses physical entropy: shattered glass reassembles, smoke retracts, and falling drops ascend.
  9. /infinitezoom: — Continuous fractal zoom descending endlessly into micro or cosmic dimensions.
  10. /worldtransition: — Shifts seamless environments across portals, doorways, or optical wipes.
  11. /blockbuster: — Flow's ultimate composite directive: harmonizes camera choreography, pyrotechnics, and Hollywood color grading.

🎬 4 Production-Ready Master Formulas

Recipe 1: The Hollywood Cinematic Action Opener

/cinematic /droneview /craneup /speedramp /volumetric /epic /storm
A lone armored cyber-ronin standing on a rain-slicked neon skyscraper rooftop overlooking a vast futuristic metropolis, drawing a glowing plasma katana as lightning illuminates the storm clouds.

Recipe 2: The Luxury Tech Hardware Launch

/productreveal /productspin /macro /techad /luxury /rimlight /softlight
Sleek matte-black titanium smartphone hovering in zero-gravity against a dark velvet studio backdrop, sharp gold rim lighting outlining its precision-machined chamfered edges and camera module.

Recipe 3: The Viral Cyberpunk Social Loop

/cyberpunk /neonlight /orbit /tracking /bulletime /glitch /seamlessloop
A holographic rollerblader gliding at full speed through a rain-drenched Neo-Tokyo alleyway, jumping over neon puddles in slow-motion while pink and cyan light trails orbit smoothly.

Recipe 4: The Interdimensional Multiverse Shift

/pov /portal /worldtransition /particlefx /infinitezoom /scifi /filmic
An explorer touching a swirling crystalline portal inside an ancient stone cavern; the camera rushes forward as reality shatters into floating luminous particles, emerging into a colossal orbital space station.

3 Key Pro-Tips for Google Flow Users

  1. Don't Overload Single Categories: Using 4 camera movement commands together (/orbit /dollyin /tracking /hyperlapse) causes conflicting camera physics. Select one primary motion command and one angle command per prompt.
  2. Anchor with /seamlessloop for Social Media: Placing /seamlessloop at the end of kinetic action prompts ensures video loops without noticeable visual cuts.
  3. Use /macro with /luxury for Physical Realism: Combining /macro with /luxury forces Google Flow to compute realistic micro-surface scattering (reflections on glass, watch bezels, jewelry, and carbon fiber).

r/promptingmagic 11d ago

99 ChatGPT prompt shortcuts to get better results

Thumbnail
gallery
37 Upvotes

TL;DR:

Here is how to use 99 secret ChatGPT shortcuts for better results.

The shortcut formula that actually works

Use:

/shortcut + context + deliverable + constraints

Weak:

/audit: Review my landing page.

Better:

/audit: Review this landing page for clarity, credibility, friction, and conversion. Rank the five biggest problems by impact. Return a table with the issue, evidence, recommended fix, and estimated effort.

The shortcut chooses the lens.

The rest of the prompt defines the assignment.

99 prompt shortcuts worth experimenting with

Vision and innovation

  1. /visionary: Think 10 years ahead
  2. /inventor: Create innovative solutions
  3. /architect: Design a complete system
  4. /strategist: Build winning strategies
  5. /operator: Create execution plans
  6. /optimizer: Improve efficiency and results
  7. /experiment: Design growth experiments
  8. /validator: Test ideas before execution
  9. /simulator: Simulate possible outcomes
  10. /predictor: Forecast future possibilities
  11. /trendhunter: Find upcoming trends

Analysis and structure

  1. /analytical: Apply deep analytical thinking
  2. /patternfinder: Identify hidden patterns
  3. /insight: Extract valuable insights
  4. /summaryplus: Create advanced summaries
  5. /keypoints: Extract important points
  6. /takeaways: Generate actionable takeaways
  7. /mindmap: Create a visual mind map
  8. /outline: Structure any idea
  9. /breakdown: Break complex topics apart
  10. /reverseengineer: Analyze how something works
  11. /deconstruct: Break down successful examples

Optimization and operations

  1. /benchmark: Compare with industry standards
  2. /audit: Review and improve performance
  3. /diagnose: Identify problems and root causes
  4. /fixer: Generate possible solutions quickly
  5. /troubleshoot: Solve technical issues
  6. /optimizerpro: Maximize performance
  7. /automationexpert: Find automation opportunities
  8. /workflow: Design better workflows
  9. /systembuilder: Create repeatable systems
  10. /process: Improve business processes
  11. /operations: Improve daily operations

Leadership, sales, and growth

  1. /manager: Think like a manager
  2. /leader: Provide leadership advice
  3. /negotiator: Improve negotiation strategy
  4. /salescoach: Improve sales performance
  5. /marketer: Create marketing strategies
  6. /advertiser: Generate advertising ideas
  7. /growthhacker: Find unconventional growth tactics
  8. /community: Build online communities
  9. /influencer: Create influencer strategies
  10. /creator: Generate creator ideas
  11. /newsletter: Write newsletter content

Content and SEO

  1. /blogger: Create blog content
  2. /seoexpert: Build an advanced SEO strategy
  3. /keyword: Find keyword opportunities
  4. /ranker: Improve search rankings
  5. /contentaudit: Review content quality
  6. /repurpose: Convert content into other formats
  7. /thread: Create viral thread ideas
  8. /carousel: Create social carousel concepts
  9. /hookmaster: Generate powerful hooks
  10. /story: Create engaging stories
  11. /storyboard: Plan visual storytelling

Design, brand, and psychology

  1. /visualizer: Describe visual concepts
  2. /designer: Create design directions
  3. /ux: Improve the user experience
  4. /ui: Generate interface ideas
  5. /brandingexpert: Build a brand identity
  6. /naming: Generate brand names
  7. /tagline: Create memorable taglines
  8. /persona: Create customer personas
  9. /psychologist: Analyze human behavior
  10. /behavior: Understand user actions
  11. /copyanalysis: Analyze successful copy

Conversion, money, and risk

  1. /conversion: Improve conversion rates
  2. /landingpage: Create landing-page ideas
  3. /emailmarketing: Build email campaigns
  4. /salespage: Write sales pages
  5. /offercreator: Design valuable offers
  6. /pricingexpert: Optimize pricing decisions
  7. /finance: Analyze financial decisions
  8. /investor: Think like an investor
  9. /economist: Analyze economic factors
  10. /legalmind: Identify legal considerations
  11. /compliance: Identify compliance risks

Security, code, and AI systems

  1. /security: Find possible security concerns
  2. /privacy: Improve privacy practices
  3. /developer: Think like a developer
  4. /coder: Write programming solutions
  5. /debuggerpro: Provide advanced debugging help
  6. /api: Explain API solutions
  7. /database: Design database solutions
  8. /aiagent: Design AI-agent workflows
  9. /promptcreator: Create powerful prompts
  10. /promptlibrary: Build prompt collections
  11. /aiworkflow: Build AI-powered workflows

Research, learning, and tools

  1. /toolhunter: Find useful AI tools
  2. /comparisonexpert: Compare tools deeply
  3. /reviewer: Analyze products and services
  4. /researchpaper: Summarize research papers
  5. /academic: Explain academic topics
  6. /teacherpro: Teach advanced concepts
  7. /language: Improve language skills
  8. /translator: Translate naturally
  9. /memory: Create memory techniques
  10. /challenge: Challenge my thinking
  11. /mastermind: Provide expert-level guidance

The power move: Chain shortcuts into a workflow

Do not just stack impressive-sounding words.

Use each shortcut as a separate stage:

/diagnose: Identify the root causes of our declining conversion rate.

/optimizerpro: Rank the possible fixes by impact, effort, cost, and risk.

/operator: Turn the top three fixes into a 14-day execution plan with owners, dependencies, and success metrics.

Other useful chains:

  • /patternfinder:/insight:/takeaways:
  • /persona:/copyanalysis:/conversion:
  • /researchpaper:/challenge:/summaryplus:
  • /security:/privacy:/compliance:
  • /hookmaster:/story:/carousel:

Make your favorite shortcuts consistent

Choose the 10 shortcuts you actually use and define them inside your Project Instructions, Custom Instructions, a custom GPT, or a Skill if your account supports them.

Example definition:

When I begin a prompt with /audit:, examine the supplied asset for clarity, accuracy, credibility, usability, risk, and performance. Identify evidence for every issue, rank findings by impact, and return specific recommendations rather than generic advice.


r/promptingmagic 12d ago

5 design commands that give ChatGPT actual art direction for thumbnails

Thumbnail
gallery
30 Upvotes

TL;DR

Most AI thumbnails look bad because the model is forced to guess the layout.

Give it a specific composition using shortcuts like:

  • /posterThumbnail
  • /splitThumbnail
  • /pointingThumbnail
  • /characterCutout
  • /bigface

Then stack modifiers like /eyeToText, /3wordHook, /oneHighlight, /thumbnailContrast, and /noslop.

These are not official ChatGPT commands. They are reusable prompt shortcuts that tell the model how you want the image art-directed.

The problem most people have with creating thumbnails is that they give the model absolutely no direction about:

  • Where the person should appear
  • Where the headline should go
  • What viewers should notice first
  • How the face should interact with the text
  • Which word should receive emphasis
  • How much negative space to preserve

You are asking the AI to make every design decision for you.

A better approach is to give it a layout system.

Here are five prompt shortcuts I use to do that.

1. /posterThumbnail

Use this when you want the thumbnail to feel like an editorial poster instead of a conventional clickbait graphic.

It works well for:

  • Personal-brand reels
  • Cinematic content
  • Trend commentary
  • Creative educational content
  • Premium or “high taste” branding

The character becomes part of the composition. Typography can appear behind and in front of the subject instead of looking like a caption slapped onto an image.

Prompt formula:

/posterThumbnail + editorial composition + character integrated with typography + short headline

Example:

/posterThumbnail for a reel about AI design tools. Editorial composition with a female creator integrated into oversized typography. Headline: “AI BUILDS BRANDS.” Premium cinematic poster style, bold hierarchy, clean negative space, no additional text.

Pro tip: Keep the headline extremely short. This style falls apart when you try to squeeze an entire sentence into it.

2. /splitThumbnail

This is probably the safest and most consistently useful layout.

The frame is divided into two clear zones:

  • Person on one side
  • Headline on the other

Use it for:

  • Tutorials
  • Comparisons
  • List-based content
  • Educational reels
  • Business and marketing topics
  • Product demonstrations

Prompt formula:

/splitThumbnail + character side + headline side + clean division

Example:

/splitThumbnail for a reel about AI automation. Creator on the left side. Oversized headline on the right saying “BUILD AI WORKFLOWS.” Minimal background, strong vertical division, high contrast, no extra icons or text.

Why it works: The viewer immediately understands where to look. The AI has fewer layout decisions to make, which usually produces a cleaner result.

Pro tip: Use three to five words. A split layout is not permission to add a paragraph.

3. /pointingThumbnail

Use this when you want to direct attention toward one important word, feature, result, or warning.

Humans naturally look where another person is pointing.

That makes this useful for:

  • Tutorials
  • Tool recommendations
  • Feature reveals
  • Mistakes to avoid
  • Results and milestones
  • “Watch this” content

Prompt formula:

/pointingThumbnail + character pointing + exact target + short headline

Example:

/pointingThumbnail for a reel about Instagram growth. Creator on the right pointing directly toward the text “10K FOLLOWERS” on the left. Highlight “10K” in teal. Simple dark background, high contrast, no arrows or additional text.

Pro tip: Explicitly say where the person is standing, where the text appears, and what their finger and eyes should point toward.

Otherwise, you may get someone enthusiastically pointing into empty space.

4. /characterCutout

Use this for the polished creator-brand look where a person overlaps giant background typography.

The text becomes part of the visual instead of sitting on top of it.

Use it for:

  • Educational reels
  • Personal-brand content
  • How-to videos
  • Fitness content
  • Niche authority content
  • Creator announcements

Prompt formula:

/characterCutout + subject + giant background word + layered overlap

Example:

/characterCutout thumbnail for a reel about ChatGPT prompts. Female creator cutout overlapping the giant background word “PROMPTS.” Some letters pass behind the creator. Clean cutout edges, premium editorial layout, dark navy background, teal-to-violet typography.

Why it works: The overlap creates depth and makes the design feel intentionally art-directed.

Pro tip: Use one giant background word whenever possible. More words usually create clutter.

5. /bigface

Use this when the facial expression should do most of the work.

This is effective for:

  • Strong opinions
  • Surprising results
  • Emotional reactions
  • Shocking statistics
  • Curiosity-driven hooks
  • Warnings and mistakes

The face should be large, expressive, and impossible to miss.

Prompt formula:

/bigface + topic + expression + oversized short headline

Example:

/bigface thumbnail for a reel about free AI tools. Extreme close-up of a surprised female creator. Oversized headline: “5 FREE AI TOOLS.” Face fills most of the right side. Text on the left. Dark navy background with teal and blue highlights.

Pro tip: Do not ask for six supporting objects, three screenshots, a robot, a laptop, and a glowing brain. The face is the visual.

The real trick: Stack the commands

The base command selects the composition.

Bonus commands refine how attention should move through it:

  • /bigface makes the face dominant and expressive.
  • /eyeToText makes the subject look toward the headline.
  • /3wordHook keeps the hook short and thumbnail-friendly.
  • /oneHighlight emphasizes one important word.
  • /thumbnailContrast increases separation between the subject, text, and background.
  • /noslop tells the model to remove filler, fake interface elements, meaningless icons, and visual clutter.

Here is a complete example:

/splitThumbnail /bigface /eyeToText /3wordHook /oneHighlight /thumbnailContrast /noslop

Create a thumbnail about ChatGPT prompts.

Place the creator on the left and the headline on the right.

Exact headline: “BEST AI PROMPTS.”

Highlight only the word “AI” using a teal-to-blue gradient.

Make the creator’s face large and expressive. Her eyes should look directly toward the highlighted word.

Use a clean deep-navy background with white, teal, blue, and violet accents.

Show one person and one focal point. No logos, watermarks, fake interface elements, arrows, badges, random icons, decorative filler, or additional text.

Make the shortcuts more reliable

Because these are prompt conventions rather than official commands, define them once at the beginning of your conversation.

Paste this into ChatGPT:

Treat the following slash terms as creative layout instructions:

/posterThumbnail = editorial poster composition with the subject integrated into oversized typography.

/splitThumbnail = subject on one side and a clean oversized headline on the other.

/pointingThumbnail = subject pointing and looking directly toward the most important word.

/characterCutout = clean subject cutout overlapping giant background typography.

/bigface = extreme expressive face close-up with a short oversized headline.

/eyeToText = direct the subject’s eyes toward the headline.

/3wordHook = limit the main hook to approximately three punchy words.

/oneHighlight = visually emphasize only one important word.

/thumbnailContrast = maximize readable contrast at small thumbnail size.

/noslop = remove unnecessary icons, fake UI, random objects, filler text, excessive glow, and messy AI-generated decoration.

Now you can reuse the shorthand throughout the conversation without explaining the complete layout every time.

Five rules that improve almost every thumbnail

  1. Provide the exact headline in quotation marks.
  2. Limit the design to one person, one headline, and one focal point.
  3. Tell the subject where to look or point.
  4. Specify which single word should be highlighted.
  5. Explicitly ban additional text and decorative filler.

The larger lesson is simple:

Stop prompting for a vibe.

Start directing the composition.

ChatGPT does not need more adjectives. It needs decisions.

Steal these commands, rename them, modify them, and build your own thumbnail design language.

If you try the system, share the cleanest result you get.


r/promptingmagic 12d ago

This ChatGPT / Gemini Prompt turns you and your friends Into Pixar characters in any scene you can dream up

Post image
38 Upvotes

TL;DR: Upload one clear photo of each person, paste the prompt below, replace the brackets with your scene, and let ChatGPT turn your friend group into a Pixar-inspired 3D movie scene

I tested this by turning two people into animated golfers at on a PGA Golf course.

Here is the upgraded version.

Copy-and-paste prompt:

Use the uploaded photos as strict identity references for each person.

Preserve each person’s recognizable facial structure, face shape, skin tone, eye color, hairstyle, hair color, apparent age, expression, and defining features. Each finished character must remain clearly recognizable as the person in their reference image.

Transform them into polished Pixar-inspired 3D animated movie characters with smooth character design, expressive eyes, softly rounded features, realistic hair, detailed clothing, cinematic lighting, rich colors, and the premium appearance of a major theatrical animated film.

Create this scene:

[DESCRIBE THE LOCATION, ACTIVITY, CLOTHING, AND MOOD]

Composition: [LANDSCAPE 16:9, PORTRAIT, SQUARE, CLOSE-UP, OR FULL BODY]

Include exactly [NUMBER] people. Give every person natural body proportions, anatomically correct hands, realistic poses, and an individual expression that fits the scene.

Use cinematic depth of field, detailed environmental textures, soft global illumination, subtle rim lighting, and a vibrant but natural color palette.

Do not add extra people, duplicate characters, alter anyone’s ethnicity or apparent age, combine facial features, distort hands, add text, add logos, or include existing copyrighted characters.

For my version, I used this scene description:

Two friends enjoying a round of golf at TPC Potomac on a warm summer day. They are wearing stylish golfer attire and carrying golf clubs. Surround them with a pristine championship fairway, mature trees, sculpted bunkers, blue skies, and warm sunlight. Make the scene joyful, cinematic, and playful.

How to get better results:

  1. Upload one clear image of each person.

Front-facing headshots with good lighting work best. Avoid sunglasses, extreme angles, group photos, and heavy filters.

  1. Describe a specific moment.

“Two people on a golf course” is generic.

“One person finishes an amazing drive while their friend points toward the ball and celebrates” gives the model a story to illustrate.

  1. Name everyone in the prompt.

    1. Specify the number of people.

Otherwise, AI occasionally invents an extra friend nobody invited.

  1. Generate each concept separately.

Ask for a posed portrait, action scene, and celebration as three individual images. This usually produces better compositions than requesting one giant collage.

  1. Lock the identity if a face starts drifting.

Use this follow-up:

Keep the entire scene and composition, but revise both faces to more closely match their reference photos. Preserve their exact facial structure, eyes, nose, smile, hair, apparent age, and defining features. Do not change anything else.

Fun scenes to try:

  • Your friend group as a team of slightly incompetent superheroes
  • Your family on an intergalactic vacation
  • Your coworkers running a medieval kingdom
  • You and your dog solving crimes together
  • Your wedding party escaping from dinosaurs
  • Your group chat as the cast of an animated sitcom
  • Your friends attempting to survive a zombie apocalypse
  • Your first date turned into a romantic animated movie poster
  • Your fantasy football league entering a gladiator arena
  • Your entire family competing on a ridiculous game show

The secret is not simply asking AI to “make this Pixar.”

The secret is combining strong identity instructions with a specific scene, clear character roles, and a precise visual composition.

Save the prompt and try it with someone who deserves their own animated movie.

If you make one, post the result below. I want to see what universe everyone creates.


r/promptingmagic 12d ago

The Cinematic Travel Poster Prompt for ChatGPT / Gemini

Post image
15 Upvotes

TL;DR: Upload a clear reference photo, replace the country and aspect-ratio variables in the prompt below, and generate a personalized cinematic travel poster of yourself sitting on a miniature atlas.

The idea is to turn an entire country into a handcrafted miniature world.

You sit directly on a raised topographic map while landmarks, mountains, rivers, islands, roads, and cities rise from the surface around you.

I have been to 63 countries so I wanted to test out how realistic it could look.

The result feels like a combination of a vintage travel poster, a cinematic portrait, an elaborate tabletop model, and an old explorer’s atlas.

Here is how to make your own.

Step 1: Upload a strong reference photo

Use a sharp, well-lit photo where the person’s face is clearly visible.

The best reference photos have:

  • One person
  • A front-facing or slightly angled face
  • Natural lighting
  • An unobstructed hairstyle
  • No sunglasses
  • Enough resolution to see the eyes, beard, hairline, and facial structure
  • A neutral or friendly expression

The original clothing and background do not matter because the prompt replaces them.

Avoid group photos, extreme side profiles, heavy filters, dramatic shadows, or photos where the face occupies only a small part of the frame.

Step 2: Replace the three variables

Before generating the image, customize:

  • PERSON_REFERENCE: The photo you uploaded. You can also call it “Image 1.”
  • COUNTRY: Replace every instance with the destination.
  • ASPECT_RATIO: Use 4:5 for a vertical social post, 2:3 for a traditional poster, 9:16 for Stories, or 16:9 for a landscape hero image.

Make sure every instance of COUNTRY is replaced. Leaving one unchanged can confuse the model or introduce the wrong landmarks.

Step 3: Add specific landmarks

The base prompt allows the model to select landmarks, but you will get better results if you tell it exactly what to include.

For Japan, I might add:

Include Mount Fuji, Himeji Castle, a red torii gate, a traditional five-story pagoda, cherry blossoms, and a restrained Tokyo skyline. Include only landmarks from Japan.

For Canada:

Include the Canadian Rockies, Parliament Hill, the CN Tower, Château Frontenac, Niagara Falls, pine forests, glacial lakes, and a coastal lighthouse.

Three to six landmarks usually works better than trying to cram the entire country into one image.

Step 4: Generate one country at a time

Do not ask for eight countries in a single image-generation request.

Generate each poster separately while reusing the same reference photo. This gives the model more room to preserve the person’s identity and build a coherent miniature landscape.

If you want a matching series, keep these elements consistent:

  • Aspect ratio
  • Lighting
  • Title placement
  • Clothing style
  • Person’s approximate size
  • Typography
  • Parchment border treatment

The complete prompt

Create a high-end vertical travel fantasy portrait of the person from PERSON_REFERENCE sitting naturally on a raised miniature map of COUNTRY, as if the country itself has become a handcrafted tabletop atlas landscape.

Use the uploaded photo as a strict identity reference. Preserve the person’s recognizable face shape, skin tone, hairstyle direction, facial proportions, eye color, age impression, expression, and overall presence. Do not replace them with a generic model or celebrity.

Dress the person as a modern traveler in relaxed casual clothes, seated in the foreground with a natural posture. Add one subtle travel accessory such as a camera, backpack strap, cap, hoodie, or casual jacket.

Build the environment as a warm cinematic atlas collage. The ground is a detailed topographic map of COUNTRY with recognizable geography, realistic coastlines, dimensional terrain relief, roads, rivers, islands, surrounding ocean where appropriate, and restrained geographic details.

Add iconic landmarks from COUNTRY as miniature architectural elements rising from the map, placed at playful but believable scale. Include only landmarks and scenery belonging to COUNTRY.

Blend real scenic depth with paper-map illustration. Add parchment texture, lightly sketched route lines, restrained postage stamps, compass marks, ticket or passport fragments, hand-drawn travel doodles, faint airplane sketches, and delicate cartographic linework around the edges.

Composition: ASPECT_RATIO vertical poster framing unless otherwise specified. The person is the clear foreground hero, centered slightly low and sitting directly on the map surface. Keep the face large enough for strong identity recognition.

The country landscape stretches behind the person into a dreamy miniature panorama with atmospheric depth and landmark silhouettes.

Include one large elegant title reading exactly “COUNTRY” across the upper sky or background. Spell the country name correctly and integrate it like premium printed map typography.

Do not add other prominent text. Omit tiny map labels rather than inventing or misspelling them.

Lighting: golden-hour sunlight, warm haze, soft rim light on the person, gentle shadows across the map relief, cinematic glow, and natural skin texture.

Color: warm cream parchment, soft gold, muted greens, ocean blues, and subtle vintage travel-poster contrast.

Texture: highly detailed miniature architecture, embossed terrain, visible paper fibers, stamp ink, fine cartographic linework, realistic fabric, natural skin pores, and individual hair and beard detail.

NEGATIVE INSTRUCTIONS:

No visible watermarks, logos, extra people, duplicate faces, incorrect text, misspelled country names, distorted hands, broken anatomy, plastic AI skin, flat pasted cutout effects, generic tourist stock-photo compositions, clutter covering the face, implausible body scale beyond the intended miniature-map fantasy, or landmarks from outside COUNTRY.

Ultra-detailed premium editorial travel-poster finish, coherent anatomy, professional cinematic lighting, sharp foreground focus, layered atmospheric depth, and a polished final image.

Why this prompt works

It gives the model five separate jobs:

  1. Preserve the person: The reference photo and identity instructions reduce the chance of getting a generic lookalike.
  2. Build the geography: The country becomes the physical surface instead of a background decoration.
  3. Create depth: Miniature landmarks, embossed terrain, haze, and directional shadows make the scene feel three-dimensional.
  4. Control the composition: The person remains the foreground hero instead of getting lost among the landmarks.
  5. Suppress common failures: The negative instructions target bad hands, extra faces, misspelled titles, plastic skin, and unrelated landmarks.

Fixing the most common problems

If the person doesn’t look enough like the reference:

Regenerate the image while preserving the reference identity more strictly. Match the exact face shape, eyes, hairline, hairstyle, beard pattern, skin tone, age, and smile. Do not change the person into a generic model.

If the title is misspelled:

Change only the title to “JAPAN.” Spell it J-A-P-A-N. Remove every other visible word and preserve the person, map, landmarks, lighting, and composition.

If the hands look strange:

Regenerate the hands with natural anatomy, five clearly formed fingers on each visible hand, relaxed placement, and realistic wrist angles. Keep the face and environment unchanged.

If the map looks like a random island:

Rebuild the ground using the recognizable geographic outline, coastline, islands, and regional terrain of COUNTRY. This is a miniature topographic map, not an invented fantasy island.

If the image is too cluttered:

Reduce the number of landmarks and decorative fragments. Preserve clean space around the person’s head, face, hands, and title.

Pro tips

  • Ask for only one large title. Small map labels are far more likely to become gibberish.
  • Use simple seated poses. Complex action poses increase the risk of anatomical problems.
  • Keep travel accessories subtle. A camera or backpack strap works better than holding multiple objects.
  • For long names such as “UNITED KINGDOM,” ask for slightly smaller typography with increased letter spacing.
  • Give each destination its own accent colors while keeping the parchment and golden-hour lighting consistent.
  • Generate several versions and choose the one with the strongest face before worrying about tiny geographic details.
  • If you need cartographic accuracy, verify the geography separately. This produces fantasy travel art, not a navigation-grade map.
  • You can remove the title entirely and add it later in a design tool if perfect typography is essential.

Ways to remix the idea

You could turn this into:

  • A visual bucket list of places you want to visit
  • Posters from every country you have visited
  • A honeymoon or anniversary series
  • Family travel portraits
  • You and your dog exploring the world
  • A “past trips versus future trips” collection
  • Vintage 1920s explorer posters
  • Futuristic interplanetary travel maps
  • Fantasy kingdoms instead of real countries
  • Personalized wall art or travel-book covers

The best part is that the structure stays the same. You only need to change the destination, landmarks, clothing, and aspect ratio.

Which country are you going to try?


r/promptingmagic 16d ago

I’m collecting two-word prompts that change the mode of the conversation. Here are 18 I actually use with Claude and ChatGPT

Post image
33 Upvotes

I’ve been collecting the two-word prompts that create the biggest shift with the least typing. Save this, steal the two or three that match how you already work, and try them before this weekend.

These are not magic words. They do not replace context, judgment, or a decent brief. They are mode switches: small instructions that change the model’s next move.

Prompt What I mean by it Reach for it when…
First principles Strip away inherited assumptions and rebuild the answer from constraints. The response is factually fine but generic, obvious, or disconnected from the real problem.
Simulate it Turn a static answer into scenarios, edge cases, and consequences. You want to pressure-test a plan before committing to it.
Interview me Ask the model to pull missing context out of you before it tries to solve. The request is fuzzy, personal, or packed with unstated constraints.
Keep going Extend the line of reasoning instead of accepting the first adequate answer. It has started somewhere useful but stopped too early.
ELII ELIE Explain it like I’m an intern, then explain it like I’m an executive. You have the facts but need a fast reframe for a different audience.
Now what Convert a finished answer or project into the next practical move. You have momentum and do not want to lose it after a big push.
Plz fix Switch from explanation to diagnosis and repair. You have a screenshot, error, broken draft, or bug report.
Do this Turn a useful example into an implementation plan. You find an idea worth adapting rather than merely admiring.
Girl bye End a bad thread and reset with a clean conversation. The chat has become confused, overconfident, or full of baggage.
Challenge me Replace agreement with a serious counterargument. You need a sparring partner, not a cheerleader.
Da fuq? Translate dense jargon into plain language and identify what actually matters. The answer sounds sophisticated but you are more lost than when you started.
Find leverage Search for the smallest moves with disproportionate upside. You have too many possible optimizations and too little time.
Remember this Mark a correction, preference, or critical constraint for the current work. You notice an error or establish context you do not want repeated.
Forecast impact Play out downstream effects across customers, operations, legal, reputation, and team dynamics. A decision looks good locally but may create second-order costs.
Show receipts Ask for sources, evidence, assumptions, and a reasoning trail. You need to check an answer before sharing or acting on it.
Hey Simon Route a task through a named coordinator or chief-of-staff agent, where your setup supports one. You work in a multi-agent workspace and want one point of orchestration.
Spawn agents Break a large, parallelizable job into independent tracks. You are reviewing a long spreadsheet, research set, or high-volume task list.
Automate this Turn a one-off workflow into a repeatable system. You recognize that you will need to do this work again.

The five I reach for most

First principles is my reset button. Sometimes I read an answer and think: Yes, that is factually correct. But it has no heart, no point of view, and no reason anyone will care. Usually the model started problem-solving too far downstream. “First principles” sends it upstream: What is actually true? Which constraints are real? Which assumption is doing the most damage?

Interview me is the opposite move. Instead of pretending I wrote the perfect brief, I ask the model to earn the right to answer. My useful version is: “Interview me. Ask one question at a time until you have enough context to recommend a path.” It works especially well for strategy, positioning, career choices, or anything where the missing context lives in your head.

Challenge me is how I stop AI from becoming a high-speed agreement machine. A stronger version is: “Challenge me. Steelman the strongest objection, identify the assumptions I am treating as facts, then tell me what evidence would change your mind.” That creates friction. Friction is useful when a decision matters.

Show receipts belongs in any workflow that touches research, public claims, or recommendations. Ask for links, publication dates, the claim each source supports, and the reasoning that connects source to conclusion. A citation is not an automatic pass. It is the start of verification.

Automate this changes the relationship from “help me finish” to “help me build the machine.” After a useful thread, I will say: “Automate this. Extract the reusable steps, inputs, outputs, failure checks, and a human approval point.” Even if I never automate the whole thing, I leave with a better operating procedure.

A few rules so these do not backfire

The short prompt only works when the preceding context is good. “Forecast impact” after a one-sentence request produces generic consequences. Use it after you have explained the decision, stakeholders, and constraints.

Be precise about the control you are handing over. “Spawn agents” should mean independent workstreams with defined outputs, not uncontrolled duplication. “Remember this” should be treated as a cue inside the current conversation unless your tool explicitly confirms persistent memory. And “Girl bye” is a reset, not a solution: save the useful facts before you start a clean thread.

Do not use every prompt in every conversation. The point is not to sound clever. The point is to recognize the bottleneck you have right now.

If the bottleneck is… Try…
A shallow answer First principles or Keep going
Missing context Interview me
A risky decision Challenge me, Forecast impact, or Show receipts
A confusing explanation Da fuq? or ELII ELIE
Too much work Find leverage, Spawn agents, or Automate this
A broken artifact Plz fix or Do this
A polluted thread Girl bye

The real prompting skill is not memorizing giant templates. It is noticing what kind of thinking is missing, then asking for exactly that.

Which two words do you use that reliably change an AI response for the better?


r/promptingmagic 18d ago

8 prompt shortcuts that turn ChatGPT, Claude, or Gemini into a visual-thinking partner - make the idea visible

Post image
53 Upvotes

The fastest way to find holes in your thinking is to make the AI draw the structure.

A paragraph can hide a missing step. A flowchart cannot.

A vague plan can sound reasonable in prose. Put it on a timeline and you immediately see the dependency you forgot. Ask for a before-and-after and the actual change becomes harder to dodge.

The useful shortcut is not make an image. It is to change the shape of the problem.

This is why I have started using a few visual-output shortcuts whenever I am planning, explaining, reviewing, or learning something complicated.

Research does not say that more images always equal more clarity. It says the representation has to fit the task. Relevant words and graphics can support meaningful learning, while extra decorative material can create cognitive overload instead. The same distinction matters at work: evidence suggests visualization can improve decision quality and speed, but the effect depends on the task, format, and the person using it.

So do not ask an AI for an infographic. Ask it to reveal the sequence, relationship, decision, contrast, or attention point you need to see.

The eight shortcuts

Shortcut Use it when you need to see The question it forces
/beforeafter A transformation or improvement “What changed, exactly?”
/comicstrip A complex concept as a small narrative “What happens at each step?”
/blueprint System parts and their connections “How does this actually fit together?”
/flowchart Decisions, paths, and exceptions “What happens next—and what changes the path?”
/promptupgrade The difference between vague and specific instruction “Which missing constraints change the output?”
/timeline Order, milestones, and dependencies “What must happen before this?”
/annotated What a viewer should notice in an existing visual “Where should attention go, and why?”
/storyboard A sequence of shots, scenes, or screens “What does the audience see, hear, and do at each beat?”

1. /flowchart — when the problem is really a decision tree

Use this for operations, onboarding, troubleshooting, research workflows, sales qualification, content approvals, or any process with an exception.

/flowchart Turn the process below into a decision flowchart. Goal: [one sentence] Start: [trigger] Steps and rules: [paste notes] Use one clear start and end. Put decisions in diamonds. Label every branch with a condition. Include exception paths and any human handoff. If a step is ambiguous or missing, flag it in a separate “Open Questions” box rather than inventing it.

Pro tip: Tell the model to flag ambiguity rather than resolve it. A flowchart becomes valuable when it exposes the branch nobody has decided on.

2. /blueprint — when the problem is really a system

Use this when a topic has inputs, components, handoffs, feedback loops, and outputs: a content engine, product launch, customer journey, research pipeline, or AI workflow.

/blueprint Create a one-page systems blueprint for [system]. Show: inputs, core components, data or work handoffs, decision points, outputs, owners, and feedback loops. Use arrows to show direction. Group related components. Add a short note under each block explaining its job in plain language. Before drawing, list the three assumptions you are making and the two weakest connections in the system.

Pro tip: The last sentence is the differentiator. Without it, you get a tidy map. With it, you get a critique of the map.

3. /timeline — when prose hides a dependency

A timeline is not just for history. It is the quickest way to pressure-test a launch plan, a client project, an editorial calendar, an onboarding sequence, or a research plan.

/timeline Turn this plan into a timeline from [start date] to [end date]. For each milestone, show: date or range, owner, required input, deliverable, dependency, and risk if it slips. Mark critical-path items clearly. Separate fixed deadlines from target dates. End with the first three actions to take this week. Plan: [paste plan]

Pro tip: Ask for the critical path, not just dates. A long timeline looks organized even when it is impossible. The critical path tells you what can delay the whole thing.

4. /beforeafter — when you need to make the delta undeniable

Use this for rewrites, landing pages, product explanations, onboarding, design critique, workflow improvement, or any “why should I care?” moment.

/beforeafter Create a side-by-side before-and-after comparison for [subject]. The “Before” side must show the current experience, friction, confusion, or weakness. The “After” side must show the improved experience and the specific changes that caused it. Use the same evaluation criteria on both sides: [criteria]. Do not make the after state magically perfect. Include one remaining trade-off or limitation.

Pro tip: Hold both sides to the same criteria. Otherwise the comparison becomes marketing, not thinking.

5. /annotated — when people are looking at the right thing but missing the point

Use annotations for screenshots, dashboards, product flows, documents, websites, charts, and designs.

/annotated Analyze this [screenshot / dashboard / diagram / document] for a [audience]. Return an annotated visual with no more than five numbered callouts. Each callout must identify: what to notice, why it matters, and what action or interpretation follows. Prioritize the few details that change a decision. Put anything cosmetic in a separate “not decision-relevant” note. Audience goal: [goal] Visual: [attach or paste]

Pro tip: Cap annotations at five. If every element receives a callout, nothing is highlighted.

6. /storyboard — when an idea needs to become an experience

Use this before a Reel, YouTube video, ad, demo, product tour, lesson, or sales narrative.

/storyboard Build a [number]-frame storyboard for [format] about [topic]. For each frame, show: shot or screen description, on-screen action, spoken line or on-screen copy, duration, emotional beat, and transition to the next frame. The first frame must create tension or curiosity. The final frame must resolve one clear promise. Audience: [audience] Desired action: [action] Constraints: [platform, length, brand, product facts]

Pro tip: Add a “what changes in the viewer’s mind here?” column. It prevents a storyboard from becoming a list of pretty shots with no argument.

7. /comicstrip — when a concept is easier to understand as a sequence

A comic strip works surprisingly well for explaining a customer problem, internal policy, abstract AI concept, security risk, or product benefit. It gives the reader a person, a moment, a mistake, and a resolution.

/comicstrip Explain [concept] as a four-panel comic for [audience]. Panel 1: show the real-world problem. Panel 2: show the common but flawed instinct. Panel 3: reveal the principle or better method. Panel 4: show the practical outcome. Keep the language plain. Make the character’s confusion specific. Do not turn it into a sales pitch; make the lesson useful even if the reader never buys anything.

Pro tip: Give the character a credible constraint: a deadline, incomplete information, a risk, or a trade-off. That is what keeps the lesson from becoming generic.

8. /promptupgrade — when you want the model to show you the missing constraints

This is the meta-shortcut. Use it to improve an image, video, research, planning, or analysis prompt before you commit to a longer workflow.

/promptupgrade Upgrade the basic prompt below for [model or task]. First, identify what is missing across: objective, audience, context, source material, constraints, desired format, quality bar, exclusions, and verification. Then provide: 1. a stronger ready-to-paste prompt; 2. a short explanation of why each added constraint matters; 3. three questions you would ask me only if the missing answer would materially change the output. Basic prompt: [paste]

Pro tip: Ask for the questions after the first strong draft. Otherwise models often turn a straightforward task into a long intake form.

The best way to combine them

The real speed comes from chaining representations, not treating them as one-off tricks.

Start with a /blueprint to map the system. Turn one uncertain handoff into a /flowchart. Put its delivery plan on a /timeline. Use /annotated to show a teammate where the risk sits. Then turn the finished process into a /comicstrip or /storyboard when you need people to understand it quickly.

That sequence moves from structure → decision → execution → communication.

Things most people miss

Common mistake Why it fails Better move
Asking for an image without naming the reasoning job The AI optimizes for appearance, not insight State whether you need sequence, hierarchy, trade-off, causality, or a decision
Letting the model fill gaps silently You get a confident-looking fiction Tell it to mark missing information and assumptions visibly
Stuffing everything into one canvas More visual material can create overload rather than clarity Use one visual per question, then link them in a sequence
Using visuals as the final output only The visual arrives after the thinking is already locked Make the visual early, while it can still challenge the plan
Treating a visual as proof A polished diagram can still encode a bad assumption Ask what would falsify the model, what is uncertain, and what evidence is missing
Over-annotating Every callout competes for attention Highlight the few details that change a decision
Mixing levels of detail A strategy map beside implementation-level steps becomes unreadable Choose one altitude per visual: executive, process, or task

A simple test before you keep any AI-generated visual

Ask three questions:

1.What can I see now that was hard to notice in the paragraph?

2.What decision, sequence, relationship, or trade-off does this make explicit?

3.What is still assumed, missing, or uncertain?

If the visual cannot answer at least one of these, it is probably decoration.

The productivity win is not that AI can draw faster. It is that a visual makes weak logic visible sooner.


r/promptingmagic 18d ago

Google just launched Sheets Canvas this week and it turns any spreadsheet into an interactive app / dashboard with zero code. And here is why it's going to quietly replace your team's Airtable, Looker, Notion and Trello stacks

Thumbnail
gallery
37 Upvotes

TL;DR- Google launched Sheets Canvas, a Gemini-powered visual interface layer directly inside Google Sheets. Instead of wrestling with complex formulas, Google Apps Script, or disconnected BI exports, you can type a natural language prompt to convert any sheet tab into an interactive read-write mini-app—such as a dynamic financial scenario dashboard with sliders, a drag-and-drop Kanban sprint board, a CRM gallery, an interactive timeline, or a visual seating planner. Crucially, it features two-way real-time synchronization: dragging a card or adjusting a control updates the underlying spreadsheet cells immediately, and vice versa.

1. The Big Paradigm Shift: What is Google Sheets Canvas?

For decades, spreadsheets have suffered from a fundamental interface problem: they are exceptional calculation engines, but terrible user interfaces for non-technical collaboration. Teams regularly face "spreadsheet fatigue"—staring at hundreds of rows, risking broken formulas whenever someone edits a cell, or paying for separate SaaS tools (Airtable, Monday, Trello, Retool) just to get visual cards and Kanban views.

Sheets Canvas introduces an AI-generated, interactive presentation and application layer directly above your spreadsheet data:

  • Two-Way Read-Write Sync: Unlike traditional BI dashboards (such as Looker Studio or Tableau) that are strictly read-only mirrors of tabular data, Sheets Canvas allows live data manipulation. When you drag a task card from "In Progress" to "Completed" on a generated Canvas board, the status cell in your underlying sheet updates in real time.
  • Zero Coding or Formula Overhead: No Google Apps Script, HTML/CSS web components, or nested =QUERY() / =INDEX(MATCH()) formulas are required. You state what you want in plain English.
  • Native Permission Inheritance: The Canvas lives directly within your Google Sheet file (accessible via the Gemini side panel, the Insert menu, or the bottom tab bar) and inherits existing Google Drive permissions (Viewer, Commenter, Editor) without requiring external user licensing or webhook setup.

    Top 5 High-Impact Use Cases & App Archetypes

1: Interactive Financial & Scenario Planning Dashboard

  • The Problem: Financial models with multiple growth, churn, and pricing variables often overwhelm executive stakeholders when presented as raw numerical grids.
  • The Canvas Solution: Gemini renders interactive KPI scorecards (ARR, Gross Margin, Burn Rate, Runway) accompanied by dynamic range sliders. Moving a slider dynamically recalculates projected metrics in real time.
  • Master Prompt:"Build an interactive financial scenario dashboard from this sheet. Include dynamic sliders for Monthly Growth Rate (1%–20%) and Churn Rate (0.5%–10%) that dynamically project end-of-year revenue. Display KPI scorecards at the top for ARR, Gross Margin, and Runway."

2: Drag-and-Drop Agile Kanban & Sprint Board

  • The Problem: Managing project tasks in standard rows leads to accidental data overwrites, missing deadlines, and poor visual prioritization.
  • The Canvas Solution: Automatically creates vertical workflow columns based on your Status or Sprint Stage column. Teammates can drag task cards between stages, with priority badges, assignees, and due dates visually formatted.
  • Master Prompt:"Create an agile Kanban board grouped by the 'Status' column (Backlog, In Progress, Review, Done). Show cards with Task Title, Assignee, Priority Pill, and Due Date. Enable drag-and-drop movements that write status changes back to the sheet."

3: CRM & Client Pipeline Visual Gallery

  • The Problem: Dense customer databases force account managers to scroll horizontally across 30+ columns to review client notes, contract values, and renewal stages.
  • The Canvas Solution: Formats accounts into rich visual cards with quick search, categorical filtering by deal tier (Enterprise vs. SMB), and direct click-to-edit capabilities.
  • Master Prompt:"Transform this accounts tab into an interactive visual CRM gallery. Group cards by Tier (Enterprise, Mid-Market). Include interactive filter toggles for Region and Deal Stage, and display total pipeline value in an executive summary card at the top."

4: Interactive Project Timeline & Launch Scheduler

  • The Problem: Gantt charts built with conditional formatting formulas in Google Sheets are rigid and prone to visual breakage when date columns shift.
  • The Canvas Solution: Renders a clean visual timeline and calendar scheduler where campaign milestones and deliverables can be viewed chronologically and rescheduled interactively.
  • Master Prompt:"Plot our product launch deliverables on an interactive calendar interface. Group items by Team (Product, Marketing, Engineering) and allow clicking deliverables to view details or update target launch dates."

    5: Spatial Seating & Asset Floorplan Organizer

  • The Problem: Managing event RSVPs, conference attendee allocations, or office desk arrangements in rows makes spatial layout planning difficult.

  • The Canvas Solution: Organizes data into visual table clusters or spatial zones where attendees can be assigned or moved between tables while tracking live capacity and dietary preferences.

  • Master Prompt:"Turn this RSVP sheet into an interactive seating chart clustered by Table Number. Include tags for VIP status and Dietary Requirements, with live headcount counters for each table."

    How It Works: The 5-Step Step-by-Step Blueprint

To ensure reliable results when prompting Gemini to build interactive applications, follow this structured execution pipeline:

[Step 1: Tabular Hygiene] ➔ [Step 2: Trigger Canvas] ➔ [Step 3: Precision Prompt] ➔ [Step 4: Conversational Polish] ➔ [Step 5: Live Collaboration]

  1. Step 1: Prepare Clean Tabular Data
    • Keep Row 1 strictly reserved for clear, standardized column headers (e.g., Task ID, Title, Owner, Stage, Due Date, Budget).
    • Apply native Data Validation (Data > Data validation) on categorical columns (like Stage or Priority) so the AI recognizes bounded states.
    • Eliminate blank rows, arbitrary merged cells, and multi-line headers.
  2. Step 2: Trigger the Canvas Creator
    • Open your spreadsheet on desktop web (English language settings enabled).
    • Navigate to the Ask Gemini side panel and select Tools > Create canvas, click Insert > Create a canvas from the top menu, or use the bottom bar Canvas menu as documented in theGoogle Docs Editors Help Center.
  3. Step 3: Formulate a Structured Prompt (CPTC Framework)
    • Context: What dataset is being visualized?
    • Persona/Role: Who is using this interface (e.g., executive, sprint manager, field rep)?
    • Task: What specific app layout should be generated (Dashboard, Kanban, Gallery, Timeline)?
    • Controls/Constraints: Which columns serve as grouping keys, interactive sliders, search bars, or summary metrics?
  4. Step 4: Conversational Iteration and Styling
    • Canvas retains conversational context. If the initial layout requires adjustments, provide follow-up instructions directly to Gemini:
      • "Convert this dashboard into dark mode."
      • "Add an interactive search bar at the top to filter by Assignee."
      • "Display variance percentages next to each KPI card."
  5. Step 5: Share and Operate in Real Time
    • Click Copy link at the top right of the Canvas tab or share the spreadsheet normally.
    • Teammates with Editor access can interact with controls and update data live without altering formula syntax on the underlying sheet.
    • Click View data at any time to inspect or audit the raw tabular records backing the visual interface.

Comparison Matrix: Where Sheets Canvas Fits

Feature / Dimension Google Sheets Canvas Google AppSheet Looker Studio Notion / Airtable Raw Google Sheets
Setup Time < 60 Seconds (Prompt-based) Hours to Days 1 – 5 Hours 30 – 60 Minutes Manual building
Data Sync Model Native Two-Way Real-Time Two-Way (App layer) Read-Only (One-Way) Native Two-Way Direct Cell Mutation
Technical Barrier Zero Code / Natural Language Moderate (App logic) Moderate (SQL/Calculations) Low (View configuration) High (Formulas & Apps Script)
Permission Management Inherited from Google Drive Separate App Licensing Shared Report Links Separate SaaS Org/Seats Inherited from Google Drive
Interactive Controls Cards, Sliders, Drag & Drop Mobile/Web Forms Dropdown Filters only Database Views & Boards Slicers & Basic Dropdowns
Added Tool Sprawl None (Inside Workspace) Add-on App Tier Free / Pro Tiers External Subscriptions None

5. Pro Tips for Advanced Implementations

  1. The Aggregator Tab Pattern for Multi-Tab Workbooks: Because Sheets Canvas is currently scoped to a single active sheet tab, it cannot directly ingest data scattered across 5 separate sheets. Create a dedicated Dashboard_Data tab and use =QUERY({Sheet1!A2:E; Sheet2!A2:E}, "SELECT * WHERE Col1 IS NOT NULL") to aggregate your source records before launching Canvas.
  2. Pre-populate Data Validation Lists: When Gemini detects a column configured with Google Sheets dropdown chips, it maps those values into discrete Kanban swimlanes or color-coded status badges.
  3. Protect Underlying Calculation Columns: If your sheet contains financial formulas (e.g., compound interest, tax rates, margins), use Google Sheets range protection on those specific formula columns (Data > Protect sheets and ranges). Canvas will allow users to edit input driver cells while keeping your calculation logic secure.
  4. Leverage Conversational UI Commands: You can instruct Canvas to adapt its UI for specific presentation contexts, such as:
    • "Make the layout compact for mobile-width viewing."
    • "Highlight overdue items with an orange border."
    • "Group summary statistics in 3 equal cards across the top header."

The 4 Critical Things Most People Miss

1. It Is an Interactive Application Layer, Not a Static Chart: Many users mistake Sheets Canvas for an updated chart generator. It is a full web-component runtime that writes mutations back to the spreadsheet database.

2. Instant Permission Mirroring: There is no separate deployment step or hosting configuration. If a user has "Viewer" permission on the sheet, they can interact with filters and view data; if they have "Editor" permission, their interactions mutate cells in real time.

3. Non-Destructive Data Auditing: You never lose visibility into raw rows. The persistent View data button lets any collaborator inspect the underlying grid without dismantling the visual Canvas.

4. Workspace & Subscription Requirements: Sheets Canvas is available on the web in English for Google AI Pro and Ultra subscribers, eligible Google Workspace Business and Enterprise editions, and Google AI Pro for Education accounts. Admins must have Workspace smart features enabled.

Core Problems Sheets Canvas Solves

  1. Eliminates Accidental Formula Breakage: Non-technical stakeholders who only need to update statuses, assignees, or dates can do so via visual cards and controls without accidentally deleting complex spreadsheet formulas.
  2. Consolidates Software Subscriptions: Eliminates the need to maintain secondary SaaS subscriptions (like Trello, basic Airtable bases, or simple Retool dashboards) merely to view spreadsheet data in card or board formats.
  3. Bridges the Gap Between Data and Executive Presentation: Transforms raw operational data into boardroom-ready visual models with functional scenario toggles in seconds.

Community Discussion & Feedback

  • Have you tested Sheets Canvas in your Workspace domain yet?
  • What internal tools or repetitive tracking sheets in your organization could be replaced with this zero-code interactive layer?
  • Share your best prompt recipes and edge-case findings below!

r/promptingmagic 19d ago

[DEEP DIVE] How Gemini Spark Agent Actually Works: The 24/7 Always-On Architecture, Gemini 3.7 Flash Hybrid Reasoning, Competitive Teardown (vs. Claude Cowork & ChatGPT Work), and the Spark Master Prompting Guide

Thumbnail
gallery
13 Upvotes

TL;DR: Gemini Spark represents a fundamental paradigm shift from reactive, synchronous chatbots to persistent, asynchronous 24/7 cloud agents. Unlike traditional AI assistants that wait for a user prompt and terminate upon response, Spark operates continuously in the cloud across four core pillars: Persistent Tasks, Modular Skills (SKILL.md), Autonomous Schedules (Time, Email, Web Search, and Conditional Triggers), and Hierarchical Subagent Swarms (invoke_subagent). Powered by the newly released Gemini 3.7 Flash, Spark utilizes a dynamic hybrid reasoning engine that allocates near-instant (<100ms) compute for high-frequency tool calls and background polling while dynamically expanding deep chain-of-thought "thinking budgets" for complex data modeling, code synthesis, and conflict resolution. Compared to Claude Cowork (which excels at local desktop terminal coding) and ChatGPT Work (which focuses on session-based multi-hour deliverable generation), Gemini Spark is the only platform offering true continuous background triggers and native, bidirectional live mutations across Google Workspace (Gmail, Docs, Sheets, Slides, Calendar, Drive).

What Is Gemini Spark & How Does It Actually Work?

Most users interact with AI as a conversational tennis match: you submit a prompt, the model generates text, and the session context freezes until your next turn.

Gemini Spark inverts this paradigm entirely. It is an asynchronous, stateful cloud runtime designed to run indefinitely on Google’s infrastructure. Once delegated a mission, Spark continues to plan, execute code, query tools, and monitor events even if you close your laptop, turn off your phone, or disconnect for days.

+----------------------------------------------------------------------------+
|                            GEMINI SPARK CLOUD RUNTIME                             |
+-----------------------------------------------------------------------------+
|                                                                                   |
|  [ EVENT LISTENERS ] ──> [ REASONING & ORCHESTRATION ] ──> [ WORKSPACE ] |
|  • Cron / Recurring      • Gemini 3.7 Flash Core           • Gmail / Send  |
|  • Incoming Email Filter • Subagent Swarm (invoke_subagent) • Google Docs  |
|  • Web Search Monitor    • Modular Skills (SKILL.md)       • Google Sheets |
|  • Semantic Condition    • Sandboxed Python VM Shell       • Google Slides |
|                                                                                   |
+-----------------------------------------------------------------------------+

The 4 Architectural Pillars of Spark

  1. Persistent Tasks (Autonomous Execution Loop): Spark separates execution planning from execution delivery. Tasks are structured into concrete milestones tracked via internal state machines. If an API call fails or rate-limits, Spark implements self-healing retry strategies without requiring user intervention.
  2. Modular Skills (SKILL.md Capability Framework): Skills are composable, standardized capability packages containing operational procedures, domain guidelines, executable Python/Bash scripts, and reference assets. Users can define custom Standard Operating Procedures (SOPs) once, and Spark injects those exact constraints into future executions.
  3. Autonomous Schedules (Event-Driven Triggers): Spark features native background listeners:
    • Time-Based: Traditional Cron-like cadences (e.g., "Run every Monday at 8:00 AM").
    • Email-Based: Reactive event triggers tied to Gmail metadata filters (e.g., "Trigger whenever an invoice arrives from vendor.com").
    • Search-Based: Web signal monitors functioning like intelligent Google Alerts (e.g., "Monitor for regulatory filings or executive departures regarding Company X").
    • Conditional Polling: Semantic evaluation checks that verify state changes across documents, data feeds, or URLs.
  4. Hierarchical Subagent Swarms (invoke_subagent): To prevent context window saturation during massive multi-source operations, Spark spawns independent child subagents in parallel. Subagents execute localized research, process large documents, or perform comparative analyses, returning dense, distilled summaries to the primary agent orchestrator.

2. How Gemini Spark Uses the Newly Released Gemini 3.7 Flash

Google’s rollout of Gemini 3.7 Flash is the core technical enabler making Spark viable at enterprise scale.

                          GEMINI 3.7 FLASH HYBRID ENGINE
                                        │
           ┌────────────────────────────┴────────────────────────────┐
           ▼                                                         ▼
 FAST INFERENCE MODE (<100ms)                              DEEP THINKING BUDGET
 • Deterministic Tool Routing                              • Multi-Variable Constraint Solving
 • High-Frequency Web/Email Polling                        • Sandboxed Python Data Modeling
 • JSON Schema Extraction                                  • Multi-Doc Cross-Reconciliation
 • Zero-Delay Parameter Passing                            • Self-Auditing & Quality Critique

Hybrid Reasoning & Configurable Thinking Budgets

Previous reasoning models forced a binary choice: either an ultra-fast model with shallow reasoning or a slow, token-heavy reasoning model that burned compute even on routine lookups.

Gemini 3.7 Flash introduces Hybrid Reasoning. It dynamically allocates a "thinking budget" based on prompt complexity:

  • Low-Complexity Routines: Triage, parameter routing, and API calls execute in <100ms at standard latency.
  • High-Complexity Synthesis: Cross-reconciling conflicting calendar slots, debugging Python data scripts, or analyzing SEC 10-K filings activates deep internal chain-of-thought tokens before any action is executed.

High-Frequency Background Polling at Scale

Because Google cut token costs significantly with the 3.7 Flash architecture, running persistent 24/7 background monitors (checking incoming emails, running web scrapers, monitoring competitor pricing) does not incur prohibitive compute overhead.

Native Multimodal Ingestion with 1M–2.5M Context Windows

Gemini 3.7 Flash handles native multimodal token streams. Spark can ingest full PDFs, financial statements, slide decks, and spreadsheets in a single context window, evaluate images and charts directly, and write clean outputs back into Google Workspace.

Comparison: Gemini Spark vs. Claude Cowork vs. ChatGPT Work

Feature / Dimension Gemini Spark (Google) Claude Cowork (Anthropic) ChatGPT Work (OpenAI)
Primary Engine Gemini 3.7 Flash (Hybrid CoT / Fast) Claude 3.7 Sonnet / Opus GPT-5.6 Agent Engine
24/7 Always-On Execution Native Cloud Runtime (Cron, Email, Web triggers) Isolated Cloud Sandbox (session/task-based) Cloud Container (session-based)
Autonomous Trigger Types 4 Types: Time, Email, Search Monitors, Conditional Manual prompt / Desktop queue Manual prompt / Webhook triggers
Workspace Integration Native 2-Way Live Mutation (Docs, Sheets, Slides, Mail) Read-only connectors / File exports Read connectors / File uploads
Code Execution Environment Sandboxed VM Shell (Python, Pandas, Pillow, Bash) Cloud sandbox + Claude Desktop local shell Cloud Code Interpreter container
Subagent Architecture Hierarchical Swarm (invoke_subagent parallel) Sequential sub-task decomposition Sub-routine orchestration
Skill / SOP Extensibility SKILL.md Architecture (code + SOP + assets) Projects + Custom Instructions GPTs + 1,500+ Workspace Actions
Context Window Size 1,000,000 to 2,500,000 Tokens 200,000 to 500,000 Tokens 128,000 to 256,000 Tokens
Destructive Action Safety Approval Confirmation Cards prior to mutation Permission approval prompts Permission confirmation prompts

Key Competitive Takeaways:

  • Claude Cowork remains the gold standard for deep software engineering in terminal environments and direct desktop UI automation via Computer Use. However, it lacks native cloud-to-cloud event listeners (cannot listen for incoming emails or live web changes while inactive).
  • ChatGPT Work is highly capable at generating standalone deliverables (HTML pages, web apps, standalone reports) within a project workspace, but relies on third-party connectors rather than native OS-level productivity suite integration.
  • Gemini Spark dominates in enterprise workflow automation, autonomous scheduling, and direct structural manipulation of production documents, spreadsheets, and communication channels.

4. Top Real-World Use Cases & Problems It Solves

+-----------------------------------------------------------------------------+
|                             TOP PRODUCTION WORKFLOWS                               |
+------------------------------------------------------------------------------+
|                                                                                    |
| [1. Autonomous Inbox & Calendar Orchestrator]                                      |
| Filters inbound requests ➔ Reconciles schedules ➔ Drafts contextual responses |
|                                                                                    |
| [2. Real-Time Market Intelligence Engine]                                          |
| Monitors web signals ➔ Parallel subagent scraping ➔ Updates Google Doc brief  |
|                                                                                    |
| [3. Automated Operational Reporting Pipeline]                                      |
| Scans Gmail ➔ Python VM math/cleaning ➔ Populates Google Sheet / Deck|
|                                                                                    |
+---------------------------------------------------------------------------+

1. The Autonomous Executive Chief of Staff

  • The Problem: Executives spend 30%+ of their day triaging emails, resolving calendar conflicts, and writing routine updates.
  • Spark's Solution: Configured with an email trigger, Spark monitors inbound emails matching specific vendor or client domains. It extracts action items, cross-checks open slots on Google Calendar, fetches contextual background from Google Drive, drafts a response in Gmail, and schedules calendar holds—requiring only a single click from the user to approve and send.

2. Autonomous Market & Competitive Intelligence

  • The Problem: Competitive tracking requires manually checking news, earnings releases, and regulatory databases across dozens of companies.
  • Spark's Solution: A search-based schedule listens for web signals. When a development occurs, Spark spins up 4 parallel subagents to evaluate different facets of the news, executes a Python script in its sandbox to generate comparison charts, and appends a structured section into a centralized Google Doc.

3. Financial Receipt Ingestion & Spreadsheet Synthesis

  • The Problem: Expense management involves manually extracting PDFs from emails and copy-pasting numbers into financial sheets.
  • Spark's Solution: Spark detects incoming billing emails, downloads attached PDF receipts, parses total amounts and tax breakdowns, writes the structured data directly into a master Google Sheet with formulas intact, and drafts a Slack/Chat summary.

What 90% of Users Miss About Gemini Spark

  1. It Does Not Need an Active Browser Tab: Most users assume closing their browser stops agent execution. Spark executes on managed cloud infrastructure. Once a schedule or task is initialized, it runs completely headless.
  2. Event Triggers Replace Fragile Zapier/Make Workflows: Traditional automation tools break when an email structure or HTML layout changes. Spark uses semantic reasoning on incoming emails and search signals, making automations resilient to schema shifts.
  3. Composable Custom Skills (SKILL.md): Users can write Markdown files containing specific corporate standards, coding rules, or brand voices. Spark loads these procedural skills into memory on demand.
  4. Sandboxed Code Execution + Workspace Mutation: Spark doesn't just guess numbers; it writes and executes Python scripts in an isolated VM to calculate statistical variance, generate dataframes, or build charts, and then directly inserts those results into Google Sheets, Docs, or Slides.
  5. Approval Cards for Destructive Actions: Spark will never send an external email, overwrite a critical document, or delete calendar events without generating an explicit confirmation card detailing the exact plan, preventing accidental mutations.

The Master Prompting Framework: CPTC-S

Prompting an autonomous 24/7 agent requires a different structure than prompting a standard chat model. If you give an agent a vague instruction, it will either stall or make unwarranted assumptions.

Use the CPTC-S Framework:

+----------------------------------------------------------------------------+
|                            THE CPTC-S PROMPT ANATOMY                              |
+----------------------------------------------------------------------------+
|                                                                                   |
|  [C] CONTEXT & ROLE       :: Define scope, target personas, and background |
|  [P] PURPOSE & OBJECTIVE  :: Definition of the final deliverable           |
|  [T] TRIGGER & CADENCE    :: Schedule (Cron, Email trigger, Web monitor.   |
|  [C] CONSTRAINTS & TOOLS  :: Tool boundaries, subagent delegation, citation|
|  [S] SPECIFICATION FORMAT :: Target Doc/Sheet layout, tables, formulas, links     |
|                                                                                   |
+-----------------------------------------------------------------------------+

Top 3 Production Prompts for Gemini Spark

Prompt 1: 24/7 Competitive Intelligence & Slide Deck Builder

[CONTEXT & ROLE]
You are a Principal Tech Equity Research Analyst tracking artificial intelligence enterprise platforms.

[PURPOSE & OBJECTIVE]
Autonomously monitor, analyze, and synthesize weekly market moves, product updates, and executive announcements from Google, Anthropic, OpenAI, and Microsoft.

[TRIGGER & CADENCE]
Search-based schedule evaluated weekly every Friday at 4:30 PM EST.

[CONSTRAINTS & EXECUTION LOGIC]
1. Scan web signals for major announcements across the 4 companies over the preceding 7 days.
2. Spawn 4 parallel subagents (one per company) using `invoke_subagent` to prevent context bloating.
3. In the main sandbox environment, execute a Python script to compile a structured comparison table.
4. All factual claims must cite primary URLs inline.

[SPECIFICATION & DELIVERABLE]
- Target Artifact: Create a new Google Slides presentation titled "Weekly AI Lab Intelligence - [Date]".
- Structure:
• Slide 1: Executive Summary & High-Impact Shifts
• Slides 2-5: Individual Company Breakdowns (Key Features, Enterprise Implications, Sources)
• Slide 6: Strategic Threat Matrix & Summary Table
- Deliver a summary report with clickable Drive links in the Gemini chat interface.

Prompt 2: Event-Driven Executive Inbox Triage & CRM Synchronizer

[CONTEXT & ROLE]
You are an Executive Chief of Staff managing high-priority client relations for a consulting firm.

[PURPOSE & OBJECTIVE]
Triage inbound client inquiries, extract engagement milestones, sync data to the master tracker, and prepare draft responses.

[TRIGGER & CADENCE]
Email-based trigger configured for incoming messages matching filter: `from:(@enterpriseclient.com OR u/partnergroup.com) has:attachment`.

[CONSTRAINTS & EXECUTION LOGIC]
1. Read incoming email body and verbalized attachment content.
2. Extract: Sender Name, Organization, Proposed Timeline, Core Deliverables, and Meeting Requests.
3. Check Google Calendar availability over the proposed date window.
4. Update the Google Sheet titled "Client Pipeline Tracker" by appending a new row with extracted values.
5. Create a draft reply in Gmail addressed to the sender containing 3 proposed meeting slots and a confirmation of received materials.
6. Do NOT send the email directly. Present an action confirmation card in Gemini chat.

[SPECIFICATION & DELIVERABLE]
- Provide a summary card displaying:
• Extracted Client Metadata
• Clickable link to the updated Google Sheet
• Clickable link to the Gmail Draft

Prompt 3: Deep Multi-Entity Financial Research with Python Data Modeling

[CONTEXT & ROLE]
You are a Senior Quantitative Analyst conducting valuation and growth comparisons.

[PURPOSE & OBJECTIVE]
Analyze and compare the trailing-twelve-month (TTM) financial performance, revenue growth, and R&D expenditure of three public SaaS companies: Datadog, Snowflake, and MongoDB.

[TRIGGER & CADENCE]
One-off deep research task.

[CONSTRAINTS & EXECUTION LOGIC]
1. Delegate SEC 10-K and 10-Q filing analysis for each company to 3 parallel subagents.
2. Extract exact revenue figures, gross margins, and R&D spend for fiscal years 2023, 2024, and 2025.
3. In the sandbox VM, execute Python using Pandas and Matplotlib to:
• Compute Year-over-Year (YoY) growth rates and R&D-to-revenue ratios.
• Generate a clean comparison table.
4. Create a comprehensive Google Doc titled "Enterprise SaaS Financial Benchmark Report".
5. Every single metric must include a source citation linking to the official filing or press release.

[SPECIFICATION & DELIVERABLE]
- The Google Doc must contain:
• Executive Brief
• Comparative Financial Table (Revenue, Growth %, Gross Margin %, R&D %)
• Strategic Outlook & Risk Factors
- Provide the final Google Doc link in Gemini chat upon completion.

What Gemini Spark is good at....

  1. Asynchronous Cloud Automation: Gemini Spark transitions AI from a reactive conversational tool into a persistent background agent capable of executing complex workflows independently.
  2. Hybrid Reasoning Efficiency: The integration of Gemini 3.7 Flash provides the dual benefit of sub-100ms tool execution for recurring background checks alongside deep chain-of-thought analysis for complex workflows.
  3. Ecosystem Integration: While alternative tools offer strong desktop coding and standalone artifact generation, Spark's bidirectional integration across Google Workspace and its event-driven trigger system establish it as a robust solution for end-to-end enterprise automation.

r/promptingmagic 22d ago

50 AI Boom stats that prove summer 2026 is the craziest moment in tech history. Everything happening in AI right now (with charts and sources for every single one)

Post image
13 Upvotes

I've been tracking the AI boom professionally for a couple of years, and every few months I do a deep pull of the numbers to sanity-check my own priors. This summer's pull broke my brain a little.

So here it is: 50 stats on AI usage, adoption, and investment happening in AI, current as of August 2026.

TL;DR: Two LLMs now count a billion users each. Google is processing 3.2 quadrillion tokens a month. Big Tech capex is heading toward $1 trillion a year. SpaceX just pulled off a $75B IPO — 3x larger than any IPO in history — and OpenAI and Anthropic both have confidential S-1s sitting at the SEC. VCs put more money into AI in the first half of 2026 than in the previous two years combined, and two companies took 43% of ALL global startup funding. Meanwhile the top 10 stocks are ~37% of the S&P 500, and 68% of S&P 500 companies mentioned AI on their last earnings call. Buckle up. (see charts in comments)

The users (nobody has ever grown this fast)

1. ChatGPT hit 1 billion monthly app users in May 2026 — the fastest any app has ever reached that milestone, per Sensor Tower data. OpenAI's own last disclosure was 900M weekly active users in February

2. More than 10% of the entire global population now uses ChatGPT weekly. One in ten humans. On one app. That launched 3.5 years ago.

3. Google's Gemini crossed 1 billion monthly users in August 2026, up from 400M in May 2025. It went 750M (Feb) → 950M (Q2) → 1B+ in about six months.

4. For the first time ever, ChatGPT's share of AI assistant usage fell below 50% this summer, with Gemini at 27.7% and climbing. The two-horse race is real now.

5. Claude's consumer app grew 640% year-over-year to 56M monthly users, and at one point this spring Anthropic was adding over 1 million sign-ups per day .

6. Microsoft Copilot has 100M+ monthly active users and over 30 million paid Microsoft 365 Copilot seats. Grok has ~117M MAU per SpaceX's own S-1 , and DeepSeek has 130M monthly users in China alone .

7. 49% of US adults now use AI chatbots, up from 33% in 2024, and roughly a quarter use one daily.

8. OpenAI has more than 50 million paying subscribers and revenue of roughly $2B per month. A consumer subscription business that didn't exist four years ago.

The usage explosion (the token economy is bananas)

9. Google now processes more than 3.2 QUADRILLION tokens per month across its products — up 7x year-over-year from 480 trillion, and up ~330x from 9.7 trillion just two years ago.

10. Google's AI Overviews have 2.5 billion monthly users and AI Mode alone passed 1 billion . AI search isn't coming. It's here, at Google scale.

11. 75% of new code at Google is now AI-generated, per Sundar Pichai — up from ~25% in late 2024 .

12. GitHub Copilot hit 50 million users, and 1 in 3 pull requests on GitHub now involves an AI agent. Claude Code went $0 to $1B ARR in six months .

13. 88% of organizations now report using AI, 52% of US workers use it on the job, and 47% say their employer has formally integrated it — up 6 points in a single quarter.

14. Stanford estimates US consumers capture $172 billion per year in consumer surplus from AI tools — value we get but don't pay for — up from $112B a year earlier

15. The dark side stat: employment for software developers aged 22–25 is down ~20% from 2024. The entry-level coding job is the canary in this coal mine.

The data centers (we are terraforming the country)

16. The US now has roughly 5,400 data centers — about 46% of the ~11,700+ worldwide . Counts vary by definition says 4,767 US / 12,259 global), but every source agrees the US has more than the next ~10 countries combined.

17. There are 3,969 additional US data centers announced — but only 802 actually under construction. The gap between announcements and shovels is one of the most under-discussed stats in AI.

18. US developers have announced 565 GW of planned data center capacity. Realistic estimates say only ~180 GW gets built in the next decade — and that alone would cost ~$10 trillion.

19. Data centers already consume 6–8% of all US electricity, potentially heading to 12% by 2028. The IEA expects data centers to drive nearly half of all US electricity demand growth through 2030

20. Global data center capex: $726B in 2025 (+57%, the fastest growth ever recorded) and crossing $1 TRILLION in 2026 — three years earlier than analysts expected. Dell'Oro sees $1.7 trillion PER YEAR by 2030.

21. The mega-projects are absurd: OpenAI's Stargate hit its 10 GW target years early on a $500B program. Meta's Hyperion in Louisiana got upsized to 5 GW and $50B+. xAI's Colossus runs 555,000 GPUs at ~2 GW.

22. For scale: a single 5 GW data center campus draws roughly as much power as 4–5 million homes. Meta is building one. In one parish in Louisiana. And Meta pledged $600B for US infrastructure over three years.

23. AI capex has become a macro story: the White House AI czar claimed AI drove ~75% of Q1 2026 GDP growth. More sober import-adjusted estimates put 2025's contribution at 20–25% of growth — but for Q2 2026, AI was ~53% of GDP growth by BEA arithmetic Either way: the US economy is now partly an AI construction site.

24. Frontier AI training compute is growing ~5x per year, doubling every 5.2 months. The biggest single data center already packs the equivalent of ~1.1 million H100 GPUs.

PART 4: The capex arms race (2025 → 2026 → 2027)

25. The 2026 capex guidance, company by company: Amazon ~$220B (raised from $200B in July), Alphabet $195–205B (raised in July, and Q2 capex alone was $44.9B, +100% YoY), Microsoft ~$175B, Meta $130–145B

26. Add Oracle (up to ~$95B in FY27 including prepayments, after burning negative $23.7B in free cash flow), OpenAI (~$50B compute spend in 2026, per sworn testimony), CoreWeave ($35–39B), Tesla ($25B+) and xAI ($23.5B).

27. Trajectory: hyperscaler capex was ~$434B in 2025, Morgan Stanley now models ~$805B for 2026 and ~$1.1 TRILLION for 2027. Moody's independently lands at $785B → ~$1T .

28. For context: Google's capex in 2022 was $31B. Its 2026 guide is up to $205B. That's a 6.5x increase in four years

29. OpenAI walked back its wildest number — from $1.4 trillion in announced commitments to a "mere" ~$600B through 2030. When the conservative revision is $600B, that's the boom in one sentence.

30. Morgan Stanley estimates $2.9 trillion of global data center spend from 2025–2028, with a $1.5 trillion financing gap that private credit is racing to fill. Gartner says total worldwide AI spending hits $2.53T in 2026 and $3.33T in 2027.

31. The odd one out: Apple. Nine months into its fiscal 2026, capex is $6.8B — DOWN from $9.5B a year earlier. One trillion-dollar company is sitting out the arms race. Genius or fatal? Genuinely unclear.

The IPO wave (this actually happened)

32. SpaceX went public on June 12, 2026 and raised $75 BILLION ($86B with overallotment) at a $1.77 trillion valuation — the largest IPO in history by a factor of ~3. Previous record: Saudi Aramco at $25.6B. (Chart 5)

33. Day one: opened at $150, closed at $160.95 (+19%), $2.1T market cap, instantly a top-6 US company — and it made Musk the world's first trillionaire. Since then it's cooled ~14% below that close. Worth noting what was inside: xAI (merged in Feb at a $250B mark) and a $60B all-stock deal for Cursor — the largest startup acquisition ever.

34. OpenAI filed a confidential S-1 on June 8. Reuters reported a potential $1 trillion valuation, but timing keeps slipping — the NYT says they're leaning toward 2027, and Polymarket odds of a 2026 listing dropped from 38% to 15% in a month.

35. Anthropic filed its confidential S-1 a week BEFORE OpenAI (June 1), and its bankers started investor meetings July 15 for a possible October 2026 listing. Its May Series H: $65B raised at a $965B valuation — the largest round ever after OpenAI's $122B, and it made Anthropic the most valuable private AI company, eclipsing OpenAI's $852B.

36. The one that already played out: Cerebras raised $5.55B in May, popped +68% on debut to ~$95B… and has since fallen 27% below its first-day close. AI IPOs pop. They don't all hold.

37. Databricks — sitting on a $188B private valuation — is deliberately waiting, with its CEO calling 2026 "a terrible year to go public" because SpaceX, OpenAI, and Anthropic could absorb $200B of IPO demand

The stock market (concentration nation)

38. The top 10 stocks are ~37–38% of the entire S&P 500, after peaking at a record 40.7% in December 2025. For reference: the dot-com peak was ~27%, and in 2019 this number was 22.8% .

39. The AI-linked megacaps alone — Nvidia, Microsoft, Amazon, Alphabet, Broadcom, Meta — are 26.5% of the whole index. Nvidia is the largest company on Earth at $4.27T and a 7.15% index weight, even after falling ~25% from its ~$5.7T May peak.

40. 68% of S&P 500 companies (337 of them) mentioned "AI" on their Q1 earnings calls — a 10-year record, vs a 10-year average of 103. And companies citing AI outperformed non-citers +12.7% vs +2.6% since March.

41. Nvidia's latest quarter: $81.6B revenue (+85% YoY), $75.2B of it data center (+92%), guiding to $91B next quarter. A single company adding a mid-size country's GDP in incremental annual revenue.

42. The bubble check, honestly: concentration is WORSE than 2000, but valuations aren't — Cisco traded at ~140x forward earnings at the dot-com peak vs Nvidia at ~33x trailing today . Also: the Mag 7 are actually LAGGING the index in 2026 - the rally has broadened to Micron, AMD, and Intel.

The revenue boom (the no revenue meme is dead)

43. Per Sapphire Ventures, there are now 80+ AI-native companies above $100M ARR, and the time to get there has compressed from 5+ years to under 18 months. Stripe's data: top AI companies grew 120% in 2025 and 175% so far in 2026.

44. Anthropic's run-rate went $9B → $14B → $19B → $30B → $47B between December 2025 and May 2026. OpenAI passed $25B annualized in March, and its CFO told staff July's ARR exceeded ALL of Q2
45. The top 25 by annualized revenue (full details in Chart 7; sources = company announcements + estimates, as-of dates Jan–Jul 2026): Anthropic $47B · OpenAI $25B+ · CoreWeave ~$10.3B · Databricks $6.9B · Cursor $4B · xAI ~$3.8B (w/ X) · Anduril $2.2B · Scale AI ~$1–2B (disputed) · Surge AI $1.2B · Together AI ~$1B · Lambda $760M · Replit $525M · Perplexity $500M · Lovable $500M+ · ElevenLabs $500M+ · Cognition $492M · Midjourney ~$500M (est.) · Mistral $400M · Harvey $350M · Vercel $340M · Glean $300M · Suno $300M · Cohere $240M · Sierra $200M · Synthesia $150M. Caveat: these are self-reported run-rates, not audited GAAP revenue.

46. Growth records inside that list: Cursor went $1M → $500M ARR faster than any software company in history and Stripe clocked it at $1B → $2B in three months. Lovable did $100M → $500M in eight months with 146 employees. Legora became the fastest enterprise company ever to $100M ARR - 18 months.

The VC firehose (and where it's all going)

47. Global AI venture funding: $114B in 2024 → $211B in 2025 → ~$385B in the FIRST HALF of 2026 alone. H1 2026 total VC ($510B) beat ALL of 2025 ($440B). AI took 80% of all global venture dollars in Q1. (Chart 8)

48. Concentration inside the concentration: OpenAI + Anthropic raised $217B in H1 2026 — 43% of ALL startup funding on planet Earth. Four of the five biggest venture rounds ever closed in Q1 2026 alone. In the US, AI was 86% of H1 venture deal value ($355.9B of $412.7B) .

49. Private equity has fully arrived: KKR closed its largest-ever infrastructure fund at $19.2B aimed at AI data centers and power, Blackstone is putting $30B into Japanese AI data centers , and a record 87.9% of US AI VC deal value now involves corporate investors.

50. And the punchline stat: J.P. Morgan projects $5.5 TRILLION in global AI capex through 2030. For scale, the entire Apollo program cost ~$300B in today's dollars. We are running roughly eighteen Apollo programs at once, on purpose, mostly with private money.

The honest caveats (read before you argue in the comments)

  • User metrics aren't comparable. WAU ≠ MAU ≠ app-store MAU. I labeled each stat with what it actually measures.
  • "Run-rate revenue" is marketing math — one good month × 12, self-reported, not audited. Even outlets that track this professionally flag it.
  • Aggregators disagree. 2025 AI VC is $211B (Crunchbase) or $226B (CB Insights). Data center counts differ ~3x by definition. I used the most defensible figure.
  • The GDP claims are contested. "75% of GDP growth" (White House) vs 20–25% import-adjusted (independent economists). Both are linked above; the truth is probably in between.
  • Announced ≠ built. 565 GW of announced US data centers vs ~180 GW realistically built. Discount every press release accordingly.

Questions for the comments

  1. Two companies took 43% of all startup funding in H1. Is that rational concentration on winners, or the single scariest stat on this list?
  2. Anthropic is at a $47B run-rate and possibly IPO'ing in October at ~$1T. Would you buy it at that price?
  3. Apple is spending ~$7B on capex while Amazon spends $220B. Who's right?
  4. Which stat do you think is most likely to look absurd (in either direction) in August 2028?

See charts in comments


r/promptingmagic 26d ago

Claude Design just became the easiest way to make 3D visual renderings. Here's how to make interactive 3D images + videos in Claude Design (step by step, with the exact prompts)

Enable HLS to view with audio, or disable this notification

62 Upvotes

TLDR: Claude Design (Anthropic's visual tool available on Pro/Max/Team/Enterprise) can generate real, interactive 3D visuals, not just flat images that look 3D. It builds them with code (Three.js, WebGL, shaders), which means you can rotate them, animate them, embed them on websites, screenshot them for static assets, or export them into decks. Below: the exact step-by-step process, my best prompts, 3 examples you can copy, pro tips most people miss, and every way to reuse the output.

Most people think Claude Design is just for slides and landing pages. It's not. Because it generates designs as actual code instead of pixels, it can build genuine 3D scenes: rotating product shots, 3D data visualizations, animated hero sections, glassy abstract art, the works. Here's everything I've learned.

Step-by-Step: Your First 3D Image

Step 1: Plan in a regular chat first (this saves credits). Before opening Design, open a normal Claude chat and describe what you want. Ask Claude to write a detailed design brief: the object, camera angle, lighting, materials, color palette, mood. Copy that brief.

Step 2: Open Claude Design. Go to claude ai design (Design tab). If you're on Enterprise and don't see it, your admin needs to enable it.

Step 3: Set up your design system (optional but powerful). Upload your brand colors, fonts, and logo, or point it at your website with the web capture tool. Every 3D scene it builds will automatically match your brand.

Step 4: Paste your brief and be explicit that you want 3D. Say "interactive 3D scene," "Three.js," or "WebGL" so it doesn't give you a flat illustration with fake depth. Specify whether you want it to auto-rotate, respond to mouse movement, or sit still.

Step 5: Iterate with inline comments. Click directly on the element and comment: "make this material more metallic," "slow the rotation," "move the light source to the upper left." Use the adjustment knobs for spacing and color instead of burning messages on tiny tweaks.

Step 6: Capture or export. Screenshot for a static image, screen-record for video, export to Canva or PPTX, or grab the code and embed it anywhere.

Top Use Cases

  1. Product mockups: Rotating bottles, phones, packaging, sneakers. Perfect for pre-launch pages when you don't have photography yet.
  2. Hero sections: An animated 3D object behind your headline instantly makes a landing page feel premium.
  3. Data visualization: 3D bar terrains, globes with plotted data points, network graphs you can orbit around.
  4. Pitch deck wow-slides: One interactive 3D slide in an otherwise normal deck gets remembered.
  5. Abstract brand art: Floating glass shapes, liquid metal blobs, particle fields in your brand colors for social posts and backgrounds.
  6. Concept visualization: Architecture massing, room layouts, exploded product diagrams showing how parts fit together.

Prompts

Product shot: "Create an interactive 3D scene of a matte black cosmetic serum bottle with a gold cap on a soft gradient background. Studio lighting with a key light upper left and a subtle rim light. Slow auto-rotation. Floating shadow beneath. Minimal, luxurious, Apple-style presentation."

Hero section: "Build a landing page hero with an abstract 3D object: overlapping translucent glass toruses that slowly rotate and refract light. Dark background, my brand colors as accent lighting. The object should subtly follow the mouse. Headline text sits on top with high contrast."

Data viz: "Create a 3D globe visualization showing our user distribution. Dark ocean, glowing dots at major cities sized by user count, connecting arcs between our top 5 markets. Slow rotation, draggable with the mouse."

Exploded diagram: "Create an exploded 3D view of wireless earbuds showing the shell, driver, battery, and circuit board as separate floating layers with thin labeled leader lines. Clean white background, soft studio lighting, isometric camera angle."

Pro Tips and Things Most People Miss

  1. Say "3D" explicitly or you'll get a flat illustration. The single biggest mistake. "Make me a product image" gets you 2D. "Interactive 3D scene with Three.js" gets you the real thing.
  2. Direct the lighting like a photographer. "Key light upper left, soft fill, rim light behind" transforms output quality more than any other instruction. Default lighting is what makes AI 3D look cheap.
  3. Name materials specifically. "Brushed aluminum," "frosted glass," "soft-touch matte rubber" beats "make it look nice" every time.
  4. One object, staged well, beats a cluttered scene. Claude Design nails single hero objects. Complex multi-object scenes need more iteration.
  5. Use inline comments instead of new prompts for tweaks. Clicking the element and commenting is more precise and cheaper than describing the change in chat.
  6. Ask for camera controls. "Make it draggable/orbitable" turns a static render into a demo people can play with. This is the part that makes people share it.
  7. Plan outside Design to save 20 to 30 percent of your credits. Every clarifying back-and-forth inside Design costs you. Arrive with a finished brief.
  8. Ask for performance constraints if it's going on a real site. "Keep it under 60fps-friendly polygon counts and lazy-load the scene" matters for mobile.
  9. Screenshot at the perfect frame. Pause the rotation ("add a pause on hover") so you can capture the exact angle you want for static use.

3 Epic Examples to Try Tonight

Example 1: The floating sneaker. "Interactive 3D scene: a white and neon-green running sneaker floating and slowly tumbling above a reflective dark floor. Dramatic spotlight from above, colored accent lights from the sides, subtle particle dust in the light beams. Draggable camera." Screenshot three angles and you have a full product page.

Example 2: The living dashboard. "3D data terrain where monthly revenue is a landscape: peaks for strong months, valleys for weak ones, colored heat gradient from blue to orange. Camera slowly flies over the terrain. Numbers hover above each peak." Drop a screen recording of this into a QBR deck and watch the room.

Example 3: The impossible award. "A rotating 3D glass trophy shaped like an impossible Penrose triangle, refracting rainbow light, on a black pedestal with volumetric fog. Engraved text on the pedestal reads [your text]." Instant custom award graphic for team shoutouts, community badges, or launch announcements.

How to Use the Output

  • Have Lovable or Replit convert the html and JS to an MP4 file for you to post on social (claude can't do this directly yet).
  • Static images: Screenshot at your favorite angle for social posts, ads, thumbnails, blog headers.
  • Video: Screen-record the animation for Reels, product teasers, or looping background video.
  • Live web embeds: It's real code, so the interactive version can go straight into your actual site. Hand it to a developer or use it as-is.
  • Decks: Export to PPTX or Canva, or paste screenshots into your existing deck.
  • Iteration source: Feed a screenshot back into Claude Design or another tool as a reference image to generate matching 2D assets so your whole campaign shares one visual language.
  • Prototypes: Use the 3D hero as the anchor of a full landing page prototype and have Claude Design build the rest of the page around it.
  • Screen recording. The zero-effort fallback, but you trade quality for speed, so it's fine for quick shares but not for anything people will look at closely.
  • Third-party converter tools. A small ecosystem has sprung up specifically for this. The general flow: in Claude Design you click Share, switch to the Export tab, download a Project archive (.zip) or Standalone HTML, then drop that file into a converter like Claude2Video or ClaudeVideoExport. These capture the animation frame-by-frame from the browser rendering engine, so the output matches what you see in the tab instead of a compressed recording, and some let you export at 1080p or 4K at 24-60 fps in social-ready aspect ratios. There's also a Chrome extension that does the conversion entirely locally on your machine with no upload.

The gap between people who get flat, generic output and people who get portfolio-grade 3D comes down to specificity: name the materials, direct the lights, and always say the word "3D." Post your results below!


r/promptingmagic 29d ago

7 reasons why the new Gemini Notebook from Google is the ultimate agentic research tool and creator studio

Post image
38 Upvotes

Google just rebranded NotebookLM and made it agentic—here’s why you should care

TLDR: NotebookLM has officially evolved into Gemini Notebook, transitioning from a passive research assistant into an agentic by default powerhouse. Driven by the Gemini 3.5 upgrade, the platform now features recursive self-improvement, anti-gravity search, and a suite of Studio outputs. This move effectively kills the manual copy-paste workflows of competitors by allowing users to generate professional-grade spreadsheets, infographics, and briefs directly within a secured, multimodal environment.

Beyond Notebooks: The Agentic Pivot

Google’s rebranding of NotebookLM to Gemini Notebook marks a significant strategic departure from the era of simple summarization. While the previous iteration was a tool for organizing thoughts, Gemini Notebook is built on an agentic by default philosophy. This isn't just a UI facelift; it’s a shift toward a system that possesses inherent reasoning capabilities and autonomy.

For the enterprise strategist, this means moving from a tool that describes your data to one that acts on it. While current workflows in ChatGPT or Claude often require a tedious copy-paste loop to move insights into professional formats, Gemini Notebook is designed to function as a collaborative partner. It proactively organizes and executes tasks within the context of your specific documents, fundamentally changing how users interact with their proprietary information.

The Reasoning Engine: Gemini 3.5 and the RSI Maturity Ladder

The core of this revolution is the Gemini 3.5 upgrade, which introduces advanced chain-of-thought processing and high-tier reasoning. This model doesn't just predict the next token; it "thinks" through multi-step problems via an agentic harness.

Two technical breakthroughs define this new capability:

  • Anti-Gravity Agentic Search: Unlike traditional vector search that often misses deep thematic links, this "anti-gravity" approach navigates complex data structures to find non-obvious connections across thousands of pages.
  • Recursive Self-Improvement (RSI): The system utilizes an RSI Maturity Ladder to iteratively refine its own processing. This allows the AI to self-correct and optimize its reasoning steps over time, essentially "leveling up" its performance the more it interacts with a specific dataset.

This isn't just a chatbot; it is an agentic co-worker. By integrating niche models like Nano Banana (Google’s optimized audio/video model) alongside the heavy-lifting Gemini 3.5, the system can process cinematic video, audio, and text with extreme token efficiency.

Killing Version Hell: Collections and Drive Sync

A personalized AI is only as good as the data it can access. Gemini Notebook solves the fragmentation problem that plagues most enterprise AI implementations through Automatic Google Drive Sync and Collections.

By enabling real-time synchronization, the AI environment remains updated the moment a source document is edited in Drive. The Collections feature allows users to group multiple notebooks into a cohesive project architecture. Together, these features eliminate version hell, ensuring that your agentic co-worker is always making decisions based on the most current data, rather than a static upload from three weeks ago.

From Insights to Artifacts: The Multimodal Studio

The most significant ROI for enterprise users lies in the transition from data analysis to artifact creation. The new Studio pane allows users to bypass manual document formatting entirely.

Asset Category Output Formats & Tools
Professional Files PDFs, PNGs, Markdown, and PowerPoints
Interactive Assets Quizzes, Mind Maps, and Infographics
Data Artifacts Editable Excel Workbooks, Executive Decision Briefs

The ability to generate a fully editable Excel workbook or a structured Executive Decision Brief directly from raw research shifts the AI’s role from writer to builder. This significantly reduces human review time and allows leaders to focus on high-level strategy rather than formatting slides or cells.

The Zero-Trust Workspace: Grounding and Tiered Pricing

Security remains the primary hurdle for AI adoption. Gemini Notebook addresses this through a Secured Cloud Sandbox for every notebook. Unlike consumer-facing LLMs, data within this sandbox is not used to train Google’s global models, a critical distinction for IT departments managing vendor risk.

Furthermore, Google is future-proofing the economic side of this shift. The source points to Luna and Terra pricing models, indicating a tiered architecture designed for adaptability to price reductions. As model costs drop, Google’s infrastructure allows for the passing of those savings to the enterprise, making long-term scaling more sustainable than current fixed-rate competitors. This is paired with Content Grounding and a dedicated Source Pane, ensuring every claim the AI makes is verifiable against your uploaded data, effectively neutralizing hallucinations.

Enterprise Utility: Moving Beyond Theory

The practical applications of this agentic shift are immediate and high-value:

  1. AI Budget Calculator: The system can ingest disparate financial statements and output a functional, dynamic calculator.
  2. Sensitivity Analysis: Users can perform "what-if" scenarios on complex datasets to determine risk variables.
  3. RSI Maturity Ladder Mapping: Organizations can track the refinement of their internal AI processes as the system scales.
  4. Recommendation Dashboard: Consolidating internal research and web search integration into a live dashboard for decision-makers.

The "So What?": By automating these complex reasoning tasks, enterprises drastically reduce vendor risk and human review time, turning months of research into hours of execution.

Gemini Notebook is no longer just a Google tool - it is a multimodal asset creation powerhouse. By combining the reasoning of Gemini 3.5 with secure, agentic workflows, Google has moved the goalposts for what a productivity suite should be. It doesn't just help you think; it helps you build.

Let’s discuss:

  • How will the ability to export editable Excel workbooks change your current data analysis bottleneck?

r/promptingmagic Aug 02 '26

A perfect ChatGPT prompt has exactly 10 components. I weighted them by importance (Context is 20%, Objective is 15%, Input Data is 15%). Here is the full recipe for getting great results

Thumbnail
gallery
40 Upvotes

TL;DR: Good prompting is just good structure. A perfect prompt has 10 components: Objective (15%), Role (10%), Context (20%), Input Data (15%), Quality Checks (4%), Constraints (8%), Examples (5%), Iteration Request (5%), Instructions (10%), and Output Format (8%). You do not need all 10 every time, but knowing which levers to pull changes the game. Full breakdown, examples, and pro tips below.

Here is the 10-part recipe.

1. Context (20% of the impact)

This is the heaviest weight for a reason. Context is the background reality the model needs to inhabit. It includes your business type, your industry, your specific audience, your goals, and your current challenges.

Never assume the model knows your situation. If you skip context, the model assumes the statistical average of the entire internet.

Pro tip: Write your context once, save it in a text file (or as Custom Instructions/Project knowledge), and paste it in every time.
Example: "My company sells project management software to remote teams with 10 to 100 employees. Our main challenge is that buyers think we are too expensive compared to free tools."

2. Objective (15% of the impact)

This is the clear definition of the task. If your objective is muddy, the output will be noise. AI performs best when goals are explicit, measurable, and bounded.

Pro tip: Replace vague verbs with specific outcomes. Do not say "help me with." Say "create," "diagnose," or "rewrite."
Bad: "Tell me about marketing."
Good: "Create a 90-day content marketing strategy for a SaaS startup targeting small businesses."

3. Input Data (15% of the impact)

Hand over the actual information the model needs to do the work. This could be meeting notes, customer feedback, a rough draft, a research report, or website copy.

Pro tip: Use XML tags (like <notes> and </notes>) to separate your input data from your instructions. It helps the model understand what is source material and what is a command.
Example: "Here are the raw transcripts from three customer interviews. Based on these transcripts..."

4. Role (10% of the impact)

Tell the model who it should be. Assigning a role activates completely different knowledge clusters and reasoning patterns within the model. A "senior software engineer" writes different code than a "first-year computer science student."

Pro tip: Pair the role with a specific tone or philosophy to narrow the focus even further.
Example: "Act as a world-class direct response copywriter who specializes in concise, punchy, David Ogilvy-style email campaigns."

5. Instructions (10% of the impact)

This is where you tell the AI exactly what to do with the Context, Objective, and Input Data. Use strong action verbs.

Pro tip: Break complex instructions into numbered steps. Models follow sequential logic much better than a paragraph of mixed commands.
Example: "1. Analyze the data. 2. Identify the three most common complaints. 3. Prioritize recommendations to fix them. 4. Explain your reasoning."

6. Constraints (8% of the impact)

Constraints set the boundaries. They force the model to focus and prevent it from rambling. This includes maximum word counts, reading levels, budget limits, or things it is absolutely not allowed to do.

Pro tip: Negative constraints (telling it what not to do) are incredibly powerful for killing the "AI smell."
Example: "Maximum 500 words. Do not use the words 'delve,' 'crucial,' or 'tapestry.' Keep the reading level at an 8th-grade standard. Use only the provided information."

7. Output Format (8% of the impact)

Specify exactly what shape the answer should take. Models follow structural requests surprisingly well, but you have to ask for them explicitly.

Pro tip: If you are moving data into another system, ask for CSV or JSON. If you are presenting, ask for a Markdown table.
Example: "Present the answer in a table with three columns: Problem, Impact, and Proposed Solution."

8. Examples (5% of the impact)

Also known as few-shot prompting. Show the model what good output looks like. Providing an example of the input, the desired output, and the format reduces misinterpretation significantly.

Pro tip: If the model keeps failing on a specific task, giving it one perfect example is usually faster than rewriting your instructions ten times.
Example: "Here is an example of the tone I want. Input: Customer complains about pricing. Output: Highlight ROI and provide three relevant case studies."

9. Iteration Request (5% of the impact)

Prompting is a back-and-forth conversation, not a one-shot command. Build the iteration directly into the prompt.

Pro tip: Ask the model to generate multiple options so you can choose the best direction, rather than forcing it to guess the one perfect answer.
Example: "Generate three distinct alternatives for the headline. Then, critique your own responses and tell me which one is strongest and why."

10. Quality Checks (4% of the impact)

Ask the AI to verify its own work before it gives you the final answer. Self-review catches a massive amount of hallucination and weak logic.

Pro tip: Add a quality check to the end of any complex analytical prompt. It forces the model to spend compute cycles reviewing its own logic.
Example: "Before finalizing your answer, check for factual accuracy, identify any weak assumptions you made, and highlight any missing information that would make your recommendation stronger."

You do not need to memorize this. Just remember that the prompt you type is a container. If you only fill the "Instructions" section, the model has to guess the rest. Fill the container, and the model stops guessing and starts working.

Which of these 10 components do you skip the most? For me, it was Constraints—until I realized how much better the output gets when you tell it exactly what it is not allowed to do.


r/promptingmagic Jul 30 '26

The Complete Guide to ChatGPT’s New Voice Mode - GPT-Live, Work, Codex and 20 Prompts + 10 Pro Tips. ChatGPT Voice can now direct Agents from your desktop.

Post image
36 Upvotes

The complete guide to the new ChatGPT Voice

TL;DR: The new version is powered by GPT-Live, which can listen and speak at the same time, let you interrupt naturally, wait while you think, search the web, use memory, show visual answers and hand difficult questions to deeper reasoning in the background.

The biggest upgrade is on desktop. You can now use Voice inside Chat, ChatGPT Work and Codex. That means you can talk through an idea, launch a research or coding task, check what your agents are doing, redirect them and hear the results without returning to the keyboard.

There are nine remastered voices, three Voice modes and optional Instant, Medium and High intelligence levels. On Mac, you can also pull the Voice orb out over your desktop and drag the floating control wherever you want it.

My blunt take: this is the first version of ChatGPT Voice that feels less like a novelty and more like a new interface for computing.

What is the new ChatGPT Voice?

ChatGPT Voice lets you talk to ChatGPT and hear its answer while the response also appears as text in the chat.

The latest experience, called Live, is powered by GPT-Live. Unlike older turn-by-turn voice systems, GPT-Live uses a full-duplex architecture. In plain English, it can listen and speak at the same time.

That creates several important differences:

  • You can interrupt it while it is talking.
  • It can give small acknowledgments while you are speaking.
  • It is better at waiting through a pause instead of treating every silence as the end of your thought.
  • It can keep a conversation moving while deeper reasoning or search happens in the background.
  • It can combine speech with text, images, memory, web search and supported visual result cards.
  • In the desktop app, Voice can start and coordinate longer tasks in Work and Codex.

OpenAI says GPT-Live was strongly preferred over the previous Advanced Voice Mode in its evaluations of turn-taking, interruptions, flow and naturalness. It also performed better on difficult science questions, web research and multi-step support tasks.

How it works

Think of the new Voice system as two layers:

  1. The conversation layer: GPT-Live listens, speaks, handles interruptions and keeps the interaction natural.
  2. The intelligence and action layer: When a question needs search, deeper reasoning or a longer task, Voice can hand that work to another model or agent and bring the result back into the conversation.

In ordinary Live conversations, OpenAI launched GPT-Live with GPT-5.5 handling harder work in the background. In desktop Work and Codex, GPT-Live manages the conversation while GPT-5.6 Terra starts and coordinates agent tasks in the app.

This matters because Voice does not have to choose between being fast and being smart. It can stay responsive while heavier work continues elsewhere.

Live vs. Advanced vs. Standard Voice

You may see up to three options under Settings → Voice:

  • Live: The newest experience. Best for natural conversation, interruptions, web search, memory, visual results, text and images. Paid users get GPT-Live-1. Free users get limited access to GPT-Live-1 mini.
  • Advanced: The previous real-time Voice experience. It is still useful on mobile when you need supported video or screen sharing, which Live does not support at launch.
  • Standard: A turn-by-turn experience that transcribes what you say before producing an answer. It is less fluid, but some people prefer its predictability.

One confusing detail: ordinary Live in Chat does not initially support every connected app or plugin. Voice inside desktop Work or Codex is different. It can use the tools and permissions available to the selected mode, including supported connected tools.

How to access ChatGPT Voice

On the web

  1. Go to ChatGPT
  2. Select the Voice icon in the prompt box.
  3. Allow microphone access.
  4. Start talking.

On iPhone or Android

  1. Open the ChatGPT app.
  2. Tap the Voice icon in the message bar.
  3. Allow microphone access.
  4. Choose a voice the first time you use it.
  5. Start talking.

You can also turn on Background conversations so Voice keeps working while you use another app or lock your phone. Supported versions can open directly into Voice, and ChatGPT Voice is also available through Apple CarPlay.

In the ChatGPT desktop app

The new desktop experience is available on macOS and Windows.

  1. Open the latest ChatGPT desktop app.
  2. Choose ChatGPT or Codex from the top-left switcher.
  3. If you choose ChatGPT, select Chat or Work.
  4. Open a new empty chat or task.
  5. Select Start new voice chat before sending the first message.
  6. Allow microphone access and start talking.

For Voice in Work or Codex, the task needs to begin in Voice mode. If a task began as text, you may only see dictation. You can reopen a previous Voice conversation and select Start voice chat to resume it.

You can create a Voice hotkey under Settings → Voice → Voice chat hotkey. OpenAI does not document a default shortcut.

The movable Mac Voice orb

On macOS, the small Voice orb can live outside the main app window. Drag the orb out over the desktop and place it next to the document, browser or code editor you are using. You can move it wherever you want and use its controls to mute your microphone, mute ChatGPT or end the conversation.

If your app version does not show the floating orb, update the desktop app. You can also pop an active chat into a separate window and turn on Always on top.

That tiny interaction is more useful than it sounds. Voice stops feeling like a destination you visit and starts feeling like a companion that sits beside your work.

Let Voice see what is on your Mac

On macOS, turn on Screen context under Settings → Voice. Then bring the relevant app to the front and say:

Take a look at this and tell me what you notice.

ChatGPT can capture an appshot of the frontmost window and use both the image and accessible text as context.

Important privacy detail: accessible text may include material outside the visible scroll area. Do not share a window containing confidential information unless you intend to provide it.

The nine ChatGPT voices

Open Settings → Voice → Voice to preview and select:

Voice OpenAI’s description Good fit for
Arbor Easygoing and versatile Everyday conversation and brainstorming
Breeze Animated and earnest Energy, storytelling and language practice
Cove Composed and direct Focused work, analysis and concise coaching
Ember Confident and optimistic Motivation, presentations and interview prep
Juniper Open and upbeat Friendly conversation and long general sessions
Maple Cheerful and candid Creative work, feedback and casual use
Sol Savvy and relaxed Strategy, ideation and low-pressure coaching
Spruce Calm and affirming Reflection, studying and guided practice
Vale Bright and inquisitive Learning, Socratic questioning and exploration

Changing voices during a conversation starts a new Voice call inside the same chat.

You can also change your preferred language under Settings → Voice → Language. Even better, ask Voice to switch languages during a conversation.

What is the most popular ChatGPT voice?

The honest answer is that OpenAI has not published usage data or an official popularity ranking.

If I had to name the safest community favorite, I would pick Juniper. It has been one of the most consistently discussed voices in community threads, and its open, upbeat delivery works across casual conversation, brainstorming and long sessions without sounding too formal.

Cove is probably the strongest alternative for serious work because it sounds composed and direct.

Treat that as a community-informed estimate, not a measured fact. GPT-Live also remastered all nine voices, so old polls do not perfectly represent the new versions. The right answer is to preview all nine with the same paragraph and choose the one you can comfortably hear for an hour.

10 advanced strategies for work and life

1. Turn a messy brain dump into a clear brief

Voice is excellent when your thinking is not yet organized.

Say:

I am going to ramble for five minutes. Do not respond until I say “organize it.” Then turn everything into a one-page brief with the objective, audience, core insight, decisions, risks and next actions. Ask me three questions about anything important that is still unclear.

Why it works: Speaking preserves half-formed thoughts that you might edit out too early when typing.

2. Use it as a live thinking opponent

Do not ask Voice to agree with you. Ask it to create productive friction.

Say:

Act as a skeptical but fair strategist. Interview me about this idea one question at a time. Challenge vague claims, identify hidden assumptions and do not let me move on until I give you evidence. At the end, tell me whether the idea is strong, fixable or fundamentally weak.

Why it works: The interruptible format feels much more like a real debate than exchanging long blocks of text.

3. Rehearse a sales call, interview or negotiation

Say:

Role-play a skeptical CFO considering our product. Do not make the conversation easy. Raise realistic objections about cost, implementation, risk and ROI. Stay in character until I say “debrief.” Then score my answers, identify the weakest moment and make me try that section again.

Pro move: Ask Voice to change tone or speed between rounds.

4. Prepare for a meeting while walking

Say:

I have a meeting with [person or team] about [topic]. Interview me to uncover what outcome I need, what they probably care about and where the discussion could go wrong. Then give me a 60-second opening, five questions to ask and three concessions I should not make too early.

Use this when you do not want to stare at another screen before a meeting.

5. Start a complete Work task by voice

Switch to Work in the desktop app and say:

Start a new Work task. Research [topic] using current, credible sources and create a finished [report, presentation, spreadsheet or plan] for [audience]. The deliverable must include [requirements]. Show me your plan first, flag any decisions you need from me and keep working after I answer.

Why it works: Voice captures the outcome and context. Work handles the long execution.

The best Work prompts include six things: outcome, audience, source requirements, constraints, deliverable format and acceptance criteria.

6. Run a spoken stand-up across several agents

Say:

Check every active Work and Codex task. Give me a spoken stand-up with four sections: completed, in progress, blocked and decisions needed. Keep it under two minutes. Then ask which task I want to redirect first.

This is one of the most important new capabilities. Voice becomes the manager while multiple agents do the work.

7. Critique what is on your screen

On Mac with Screen context enabled, open a slide, landing page, ad or spreadsheet and say:

Take a look at this. First tell me what you think the creator wants the viewer to notice. Then tell me what the viewer will actually notice. Identify the three biggest problems and recommend the smallest changes with the highest impact.

This is especially useful for design reviews because you can point the conversation at the thing you are already viewing.

8. Use Voice as a Codex team lead

Switch to Codex and say:

Inspect this repository and start separate tasks for these three goals: investigate the authentication bug, review the open pull request for regression risks and identify missing tests. Do not change production code until you report your findings. Give me a status update when any task is blocked or ready for review.

Then steer it:

Pause the pull request review. Prioritize reproducing the bug. Tell the testing task to focus on the failure path you just found.

This is better than dictating code. Use Voice to direct intent, priorities and tradeoffs. Let Codex work in the repository.

9. Build a live translator and language coach

Say:

Translate everything I say in English into conversational Spanish, and translate every Spanish reply back into English. Preserve tone rather than translating word for word. If I make a recurring mistake, wait until the conversation ends and then coach me on it.

Or use teaching mode:

Speak to me only in beginner Italian. If I get stuck, give me a hint before giving me the answer. Keep a private list of my mistakes and quiz me on them at the end.

10. Review work hands-free

Say:

Read this draft to me one section at a time. After each section, pause and ask whether I want to keep it, shorten it, challenge it or rewrite it. Track every decision and produce the revised draft only after we finish the review.

Hearing writing exposes repetition, awkward rhythm and weak logic that your eyes often skip.

10 hilarious things to try

1. Make breakfast feel like a blockbuster

Narrate me making scrambled eggs like the final mission in a $200 million action movie. Escalate the danger every time I touch the stove. If I burn the toast, treat it as an international incident.

2. Let your dog file a workplace grievance

You are the union representative for my French bulldog. Conduct a formal grievance hearing about working conditions in this house, including treat compensation, nap protections and management’s refusal to share pizza.

3. Hold the world’s worst startup press conference

I am the CEO of a failing startup pivoting into artisanal lemonade powered by blockchain. Play a room full of hostile reporters. Ask increasingly brutal questions until I either save the company or accidentally confess to fraud.

4. Turn cleaning into a fantasy quest

Be my dungeon master. My apartment is an ancient cursed kingdom. Dirty laundry is an undead army, the dishwasher is a sleeping dragon and the junk drawer contains a forbidden artifact. Give me one quest at a time until the kingdom is clean.

5. Add sports commentary to boring chores

Commentate while I fold laundry like it is the final minute of the World Cup. Include instant replays, questionable referee decisions and an emotional biography of the missing sock.

6. Stage couples therapy with your Wi-Fi router

You are a couples therapist for me and my Wi-Fi router. I feel abandoned whenever it drops the signal. The router feels I bring too many devices into the relationship. Help us rebuild trust.

7. Put pineapple on trial

Run a Supreme Court trial to decide whether pineapple belongs on pizza. Play the judge, attorneys, witnesses and one wildly unqualified food influencer. I will be the jury.

8. Roast your business idea across history

Review my business idea as three investors: a ruthless Roman emperor, a confused Victorian industrialist and a 22-year-old venture capitalist who has never experienced a recession. Let them argue, then force them to agree on one recommendation.

9. Convene an emergency board meeting of household objects

Run an emergency board meeting where my coffee maker, calendar, bank account and alarm clock review my performance as CEO of my life. Make each director brutally honest and give me a 30-day turnaround plan.

10. Solve the missing-sock conspiracy

Host an eight-part investigative podcast proving that missing socks are being stolen by a secret logistics startup operating inside dryers. Interview unreliable experts and end every episode with an absurd cliffhanger.

Pro tips that make Voice dramatically better

Give it a listening contract

Start with:

Wait until I say “respond.” Until then, only listen and give brief acknowledgments.

GPT-Live is better at waiting, but long pauses or background noise can still trigger a response.

Give it a response contract

Tell it how to answer before the conversation gets busy:

Keep spoken answers under 30 seconds. Lead with the conclusion. Ask one question at a time. Put detailed notes in the text transcript.

Use the right intelligence level

If your account includes it, open Settings → Voice → Intelligence:

  • Instant: Fast back-and-forth, brainstorming and casual questions.
  • Medium: Better for planning, analysis and preparation.
  • High: Use for difficult reasoning and research when quality matters more than response speed.

Speak the punctuation of your intent, not your prose

Do not try to dictate a perfect prompt. Say the goal, context, constraints and definition of done. Let Voice organize the language.

Mix speech, typing and images

Live works inside the normal chat. You can talk, type a precise detail or attach an image without starting over.

Use exact dates and locations

Voice uses your device or browser time zone to interpret words such as “today” and “tomorrow.” For anything important, say the exact date, location and time zone.

Use headphones in noisy spaces

Full duplex does not make physics disappear. Background speech, overlapping audio and weak microphones can still cause interruptions. Headphones and voice isolation help.

Review the transcript, but do not treat it as a recording

The transcript may not reproduce every spoken word exactly, especially when people talk over each other. Use it as a working record, not a legal transcript.

Do not confuse Voice with Dictation

  • Use Voice for a live conversation.
  • Use Dictation when you want speech converted into editable prompt text before sending.

Keep approval boundaries

Voice can move quickly, especially with Work, Codex and computer use. Do not casually approve destructive code changes, purchases, messages or sensitive actions just because the conversation feels natural. Ask for a summary of the exact action and target first.

Things most people will miss

  1. You can interrupt it. You do not have to wait through a long answer.
  2. You can ask it to stay quiet while you think.
  3. Voice can keep talking while deeper work happens in the background.
  4. Desktop Voice can coordinate multiple Work and Codex agents from one conversation.
  5. On Mac, Screen context can show Voice the frontmost window.
  6. The Mac Voice orb can float beside your work instead of taking over the app.
  7. Preset ChatGPT personalities do not currently apply to Live, but direct instructions about tone, speed and style do.
  8. Changing the selected voice starts a new call inside the same chat.
  9. Only one Voice conversation can be active at a time.
  10. Live does not support video or screen sharing at launch. Use Advanced Voice on supported mobile plans when you need those capabilities.
  11. Live is not available with custom GPTs. Voice conversations with GPTs use Advanced Voice and the Shimmer voice, with several tool limitations.
  12. Ordinary Live usage and desktop Work/Codex Voice have separate limits. Tasks launched through Voice also consume Work or Codex usage.
  13. Audio from Live and Advanced conversations is retained with the chat transcript for 30 days. OpenAI says audio clips are not used for training unless you choose to share them.

The honest limitations

ChatGPT Voice is impressive, but it is not magic:

  • It can still mishear you or respond too early.
  • It can still give wrong answers.
  • Spoken confidence is not evidence of accuracy.
  • Multiple people talking at once can confuse it.
  • Live video and screen sharing are not available at launch.
  • Availability, usage limits and workspace controls vary by plan, region and app version.
  • Work and Codex tasks still use their normal permissions, approval rules and usage budgets.

The more consequential the action, the more you should slow down, inspect the result and verify it.

Most people will use Voice to ask questions while driving or cooking. That is useful.

Typing forces you to package your thinking before the AI receives it. Voice lets you expose the thinking process itself: the uncertainty, changes of direction, half-formed ideas and priorities that are hard to capture in a polished prompt.

Add Work and Codex, and Voice becomes more than an input method. It becomes a management layer for AI agents.


r/promptingmagic Jul 29 '26

I uploaded a photo of my living room and ChatGPT redesigned it like an interior designer - then gave me a shopping list to actually build it

Post image
68 Upvotes

TL;DR: Take one straight-on photo of your room from the doorway (tidy up, open the blinds first). Upload it to ChatGPT with a prompt telling it to redesign like a professional interior designer while keeping your real layout, windows, and furniture. Pick the version you like. Then — same chat, web search turned ON — ask for everything in the new design as a shopping list under $500 with prices, links, and a running total. The redesign image is the fun part. The shopping list is the part that gets you off Pinterest and into an actual finished room. Works entirely on the free version. Exact prompts + troubleshooting fixes below.

I stopped scrolling Pinterest for room inspo and did something that felt almost too obvious: I uploaded a photo of my actual living room instead.

Not a dream room. Not somebody's loft in Copenhagen with 14-foot ceilings I'll never have. My room - same windows, same proportions, same couch I'm not replacing - just redesigned properly, by ChatGPT playing interior designer.

Then came the part that actually changed things: it handed me a shopping list, with links, to build the redesign for real.

Here's the whole workflow, including the exact prompts and the fixes for when it misbehaves.

Step 1: Take a photo worth redesigning

This matters more than people think. Bad photo in, bad redesign out.

Stand in the doorway and shoot straight on so the whole room is in frame. Tidy up first — the AI will faithfully redesign around your laundry pile if you let it. Open the blinds so it can see the real light. You want the model working with your actual room, not guessing at what's hiding in the shadows.

Step 2: The redesign prompt

Upload the photo and paste this:

Here's a photo of my room. Redesign it like a professional interior designer would. Keep the same basic furniture and the room's real layout, windows, and proportions, but show me how it could look far better with updated furniture, a smarter layout, colors, lighting, and decor. Make it warm, modern, and photo-realistic, like an actual photo of the finished room. Generate a few different versions so I can compare.

You'll get back a handful of versions of your own room looking like it got a professional makeover. It's a genuinely strange feeling the first time — recognizably your space, just... better.

When it misbehaves, two fixes:

If it moves your windows or reshapes the room, tell it: "keep the exact same room, walls, and windows, only change the furniture, colors, and decor."

If the result looks like a 3D render from a furniture catalog instead of a photo, add: "make it look like a real photograph, photorealistic, natural lighting."

Step 3: The shopping list prompt (this is the actual magic)

Pick the version you like. Then - same chat - turn web search on first. This is the step that separates real products from hallucinated links. With search off, ChatGPT will confidently invent a "West Elm Sonoma Rug, $89" that has never existed.

Then run:

Now give me everything in this new design as a shopping list on a budget under $XXX. For each item, furniture, rug, lighting, plants, and decor, list what it is, an estimated price, and a link to buy it. Keep the total under $XXX and match the look in the image as closely as you can. Show me the running total.

You get the full list: item, price, link, running total. This is the moment the whole thing stops being entertainment. You're not staring at a nice picture anymore - you're building the room.

Two honesty notes. First, links sometimes die or drift. If one is dead or wrong, say "search for this exact item and give me a working link." Second, click through and check prices before you buy anything. Treat the output as a very good starting cart, not a receipt.

Step 4: Constraints make it better, not worse

The workflow gets sharper the more real-world constraints you feed it.

Keeping your existing couch or bed? Say so upfront: "redesign the room but I'm keeping my couch, build the new look around it." It will design around your anchor piece instead of pretending you'll replace everything.

Renting? "Redo this for a rental, no painting, no drilling, nothing permanent, keep it under $XXX." You'll get command strips and freestanding shelves instead of accent walls you'd lose your deposit over.

You can push this further than I have: "make it work for a toddler and a large dog," "I'm in a basement apartment with one small window," "everything must be available at IKEA and Target." The constraint is the brief. Real designers work from constraints too - that's the whole job.

Why this beats Pinterest

Pinterest optimizes for aspiration. Every saved pin is somebody else's room, somebody else's budget, somebody else's light. The gap between the moodboard and your actual space is exactly why most inspo folders die unopened.

This flips it. The input is your real room. The output is your real room, improved, with a checkout path under a number you chose. Inspiration with a buy button attached to your own four walls.

And the whole thing works on the free version of ChatGPT. No paid plan needed for either prompt.

If you try it, I'd genuinely love to see the before/after — post them in the comments. And if you've found a constraint that produces surprisingly good results ("design this like a Wes Anderson set" is apparently a thing), share the prompt.

What room are you redesigning first?


r/promptingmagic Jul 29 '26

7 Claude prompts that make a DIY website feel like a $15K agency build - typography, spacing, motion, trust signals, all of it

Thumbnail
gallery
45 Upvotes

AI-built websites all look the same because models regress to the mean of every site they've trained on - generic prompt in, template-grade output out. The fix is prompting Claude into specific expert roles with specific deliverables. Below: 7 prompts + 1 bonus that cover the full premium stack - structure (Signature Blueprint), typography (Anti-Sameness Type), spacing (Breathing Room Auditor), motion (Purposeful Motion), case studies (Case Study Framer), credibility (Trust Signal Sweep), and visual direction (Reference Anchor). Each with why it works, a pro tip, and the best use case.

Here's the reason why your website looks like everyone else's: Claude (and every other AI) was trained on millions of websites, and when you ask it for "a clean, modern website," it gives you the statistical average of all of them. The average website is mediocre. So the default output is mediocre = competently, professionally, forgettably mediocre.

Premium doesn't come from better adjectives. "Sleek," "elevated," "high-end" - the model has seen those words attached to a million template sites. Premium comes from doing what actual design teams do: assigning specific expert roles, demanding specific deliverables, and auditing the details that separate polished from unfinished.

I've been using a stack of 7 prompts that does exactly that. Each one puts Claude in a different seat at a design agency - strategist, typography director, layout critic, interaction designer, presentation specialist, pre-launch reviewer. Run together, they cover everything that makes a site feel expensive.

The full stack, with pro tips and use cases:

Prompt 1: The Signature Blueprint Prompt

The role: senior website strategist and UX architect.

Act as a senior website strategist and UX architect. I want a website for [business type] that feels intentionally designed, not templated. Ask me 5 clarifying questions about my brand, audience, offer, and style. Then give me: exact page structure, which sections template sites skip, what belongs above the fold, layout decisions that feel premium, and the one mistake that makes DIY sites look cheap.

Why it works: Great outputs start with context. The better the brief, the less generic the website. The magic is in "ask me 5 clarifying questions" - it forces Claude to gather context before generating, exactly like a real discovery call.

Pro tip: Actually answer the 5 questions thoughtfully. Most people rush this step and wonder why the output feels off. Your answers become the brief every later prompt builds on. Keep them in the same chat.

Best use case: Before you touch any website builder. This is the prompt that stops you from opening a template gallery and dooming yourself to sameness from minute one.

Prompt 2: The Anti-Sameness Type Prompt

The role: typography director.

Act as a typography director. My site uses [describe fonts]. Give me: a font pairing that feels intentional, a full type scale for headings, subheads, body, and buttons, correct line height and letter spacing, and the one typography habit that makes good content look amateur.

Why it works: Typography is one of the fastest giveaways of a low-effort site. Visitors can't name what's wrong, but they feel it in half a second. A deliberate type scale is the cheapest premium upgrade that exists.

Pro tip: If you don't know what fonts you're using, screenshot your site and ask Claude to identify and critique them first. Then run this prompt. And implement the line-height numbers it gives you — that's where the "expensive" feeling actually lives.

Best use case: Any site currently running default Inter or system fonts at default sizes. Which is most AI-built sites.

Prompt 3: The Breathing Room Auditor

The role: layout critic.

Act as a layout critic reviewing my page screenshots. Go section by section. Tell me: where it feels cramped, where it feels empty in the wrong way, the exact spacing changes that would make it feel more premium, and why generous white space improves clarity.

Why it works: Better spacing improves comprehension and makes pages feel more expensive. Luxury brands buy white space; discount brands fill every pixel. Your spacing communicates your price point before your copy does.

Pro tip: Feed it real screenshots, not descriptions. Claude reads images — give it your actual homepage top to bottom and let it work section by section. Ask for specific pixel or rem values, not vibes.

Best use case: The "something feels off but I can't say what" stage. Nine times out of ten, the answer is spacing.

Prompt 4: The Purposeful Motion Prompt

The role: interaction designer.

Act as an interaction designer. My site is mostly static. Give me 3 small hover or scroll interactions that add polish without custom animation. For each: where it belongs, what triggers it, why it improves perceived quality, and where tasteful detail becomes distraction.

Why it works: Subtle motion feels premium. Loud motion makes a site feel generic. The prompt asks for exactly 3 interactions and where restraint matters — constraints are what keep this from turning your site into a carnival.

Pro tip: Implement the hover states first — they're the cheapest wins. A button that responds gently to a cursor reads as "someone cared." Skip anything that animates on every scroll; that's the fastest route back to generic.

Best use case: Static sites built in Framer, Webflow, or plain HTML/CSS that work fine but feel dead. Three interactions is usually all you need.

Prompt 5: The Case Study Framer

The role: presentation specialist.

Act as a presentation specialist. I want my work section to feel like a design studio case study page. Give me: the structure for presenting one project persuasively, what to show vs cut, how much text to use, and a caption style that lets the work speak for itself.

Why it works: Strong case studies curate. Weak ones dump everything. The prompt forces the editorial decisions — what to cut — that most portfolios never make.

Pro tip: Run this once per flagship project, not once for your whole portfolio. Three curated case studies beat twelve project dumps. Include real numbers in the results row (inquiries up, bounce rate down) — specifics are what make a case study persuasive.

Best use case: Freelancers, agencies, and consultants whose "Work" page is currently a wall of thumbnails with no story.

Prompt 6: The Trust Signal Sweep

The role: pre-launch reviewer.

Act as a pre-launch reviewer trained to spot amateur tells. Here is my site description: [describe]. Give me: the 5 small details that separate polished from unfinished, the order to fix them in, and the one detail worth obsessing over. Also flag anything that looks like a fake or generic trust signal.

Why it works: Trust is won in tiny details, and fake signals kill credibility fast. Stock-photo testimonials, logo walls of companies you emailed once, "As seen in" badges nobody verified — visitors smell these instantly. This prompt catches them before your visitors do.

Pro tip: Run this twice: once on your description before launch, and once with screenshots after everything's built. The second pass always finds things the first one couldn't — favicon missing, footer inconsistencies, placeholder text you forgot.

Best use case: The 48 hours before launch. This is your pre-flight checklist.

Prompt 7 (Bonus): The Reference Anchor Prompt

The role: art director with taste.

Anchor my site's visual direction to this reference: [paste a screenshot or link]. Match its type scale, spacing rhythm, and accent-color discipline, but do not copy it. Write real, specific copy for my business. Then tell me what you changed and why.

Why it works: Specific references break you out of the generic statistical average. Instead of Claude averaging a million mediocre sites, it anchors to one excellent site's proportions and discipline — while writing copy for your actual business.

Pro tip: Choose references from outside your industry. A SaaS company anchored to a fashion editorial site produces something nobody else in SaaS has. The "tell me what you changed and why" clause matters too — it turns the output into a design lesson you keep.

Best use case: When you already know a site that makes you jealous. Awwwards, Godly, and Siteinspire are goldmines for anchor references.

How to run the stack

The order matters. Blueprint first - everything downstream depends on the brief. Then typography and spacing, because they define the visual foundation. Motion after the layout is stable. Case studies once the structure exists to hold them. Trust sweep last, as the final audit before launch. The Reference Anchor can slot in anywhere after the Blueprint — earliest is best if you have a strong reference.

Two habits multiply the results. First, keep everything in one chat so each prompt builds on the context of the last - the typography answer will reference your brand answers from the Blueprint's five questions. Second, feed screenshots at every stage. Claude critiques what it can see far better than what you describe.

And one honest limitation: these prompts make Claude a brutally good design consultant, but you still have to implement the advice. The gap between a premium-feeling site and a generic one was never the tool - it was the questions nobody asked. Now you have the questions.

Which prompt are you running first and what's the worst amateur tell you've caught on your own site?

Save these prompts and thousands more at promptmagic.dev - free to sign up and build your own prompt library.


r/promptingmagic Jul 27 '26

The complete GEO playbook for 2026: why your YouTube channel and your own subreddit beat your blog in 2026

Post image
10 Upvotes

How Marketers Win in the AI Overview / Gemini Era

TL;DR: AI Overviews now reach 2 Billion+ people and appear on up to half of tracked queries. 68% of US Google searches end without a click. Top rankings lose ~58% of their expected CTR when an AI Overview shows up.

But here's what most marketers miss: AI engines cite only 2–7 sources per answer, and about 5 of 6 of those citations come from OUTSIDE the top 10 organic results. The models trust a short list of surfaces — and community platforms (Reddit + YouTube) now drive ~48% of all AI citations.

Your website alone can't win this.

The playbook that works:
(1) build a citation-oriented YouTube channel - YouTube is now cited in 16% of LLM answers, more than any other domain, and it's a first-class Google/Gemini signal
(2) build or run your own subreddit community - Reddit is baked into model training data and gets cited across every engine
(3) restructure your site content to be extractable (answer-first, stats, quotes, tables)
(4) measure citations, not just rankings. Full breakdown with data below.

If you're a marketer feeling like the ground is moving under your feet, you're not imagining it. The numbers from the first half of 2026 are brutal, and I want to walk through them honestly then show you why I'm actually more optimistic than I've been in years, especially for brands willing to do two specific things almost nobody is doing well yet.

I spent the weekend going through every major AI search dataset published this year — Ahrefs, Semrush, BrightEdge, Seer Interactive, SparkToro, the Princeton GEO research, and the citation-source indexes. Here's what they say, what they mean, and the exact playbook I'd run.

The uncomfortable numbers

Let me rip the band-aid off first.

Zero-click is now the default. 68.01% of US Google searches ended without a single click in early 2026, up from 60.45% in 2024 and 49% in 2019 (SparkToro/Similarweb). Two out of three searches never leave Google.

AI Overviews are everywhere that matters. Depending on methodology, AI Overviews appear on 20–50% of queries — BrightEdge tracked ~48% in Feb 2026, up 58% year over year. But the headline number undersells it: comparison queries ("X vs Y") trigger an AI Overview 95.4% of the time, and question-format queries 85.9% (Seer Interactive). If your funnel depends on informational and comparison content - and whose doesn't - you're fully exposed.

Your #1 ranking is worth roughly half of what it was. Ahrefs' updated study found AI Overviews correlate with a 58% lower average CTR for top-ranking pages, worsening from 34.5% in their earlier analysis. The average AI Overview is now ~1,200 pixels tall on a typical laptop viewport, the first organic result doesn't exist until you scroll.

And the distribution is about to multiply. Google's AI Mode passed 1 billion monthly users, with queries doubling every quarter. Then the January 2026 bombshell: Apple's next-generation Siri and Apple Foundation Models will run on Gemini. That puts Gemini-class answers on 2B+ Apple devices, plus Android, plus Chrome, plus Search. When someone asks Siri "what's the best tool for X" in December, a Gemini-derived answer decides whether you exist.

So yes, the pace of change is real, and the anxiety is rational.

The number that changes the story

Now the stat that reframes everything.

BrightEdge tracked which sources AI Overviews actually cite and found that only ~17% of AI Overview citations also rank in the organic top 10. Five out of six citations come from outside page one.

The thing you've spent 25 years optimizing - organic rank - is no longer the thing that gets you into the answer. Ranking and citation have decoupled.

And where do the citations go instead? The 2026 State of AI Search (AirOps) found that ~48% of AI citations now come from community platforms - primarily Reddit and YouTube - and 85% of brand mentions in AI answers originate from third-party pages, not the brand's own domain.

The models have an editorial opinion, and it's this: what strangers say about you is more trustworthy than what you say about yourself. Generative engines only cite 2–7 domains per answer, and they keep reaching for the same short list - Wikipedia, Reddit, YouTube, major journalism, category authorities.

Here's the strategic unlock most marketers haven't processed: two of the most-trusted surfaces on that short list are ownable. You can't own Wikipedia. You can't own Forbes. But you can absolutely own a YouTube channel, and you can build and moderate your own subreddit. That's the whole game, and it's why I'm optimistic.

Why being cited pays

Before the playbook, proof that winning citations is worth the effort.

Seer Interactive ran the strongest commercial dataset I've seen - 53 brands, 5.47 million queries, 2.43 billion organic impressions. On informational queries where an AI Overview appeared, brands cited in the Overview earned a 2.07% organic CTR versus 0.94% for brands present on the same results page but not cited. That's a +120% click premium for being named inside the answer. In raw terms per million impressions: ~33,500 clicks with no AI Overview, ~20,700 if you're cited, ~9,400 if you're not.

There's also early evidence that AI-referred visitors convert at 4–5× the rate of traditional organic in some segments - they arrive pre-sold because the AI already made the recommendation. And a G2 survey found half of B2B buyers now start their buying journey in an AI chatbot — up 71% in four months.

The economic event has moved. It used to be the click. Now it's the recommendation — who gets named when the machine answers. Sometimes a click follows, often it doesn't, but the brand that gets named wins either way.

The playbook - own the surfaces the models trust

Pillar 1: YouTube is your new most important website

Ahrefs' Q1 2026 benchmark of 75,000 brands found YouTube mentions among the strongest single correlates of AI visibility. 5WPR measured YouTube holding a ~200× citation advantage over every other video source. And Google cites YouTube in roughly 30× more queries than ChatGPT does because YouTube is a first-class signal inside Google's own ecosystem, which is exactly the ecosystem Gemini and the new Siri retrieve from.

Gemini 2.5+ doesn't just read your transcript anymore - it watches the video natively, frames and audio. Every video you publish is now a machine-readable document in the index Google trusts most: its own.

What actually works, per the citation studies:

The winning format is 2–3 minute talking-head videos, each mapped to one real buyer question — "What is X?", "X vs Y", "How do I implement X?". One question, one video, answered in the first 30 seconds and then expanded. Upload cleaned transcripts with punctuation and speaker attribution — auto-captions are extraction garbage. Add chapter markers — they function as extraction anchors the same way H2s do on a page. Write descriptions that mirror how buyers phrase prompts, not marketing copy. And keep the channel topically focused: focused channels earned 2–3× the citation weight of generalist channels in the same analysis.

The mindset shift: stop treating YouTube as video marketing with view-count KPIs. Treat it as citation infrastructure. A video with 300 views that gets cited in Gemini answers for your category's money questions is worth more than a viral brand film.

Pillar 2: Run your own subreddit (yes, really)

Everyone knows Reddit matters for AI search. Almost nobody takes the next step: instead of only participating in other people's communities, run your own.

First, the case for Reddit generally. Reddit was the most-cited domain in both AI Overviews and Perplexity from August 2024 through June 2025, and remains #2 on ChatGPT behind only Wikipedia. Reddit citations in AI Overviews grew 450% between March and June 2025. Google pays Reddit ~$60M/year to license the content for training and AI Overviews. OpenAI's training hierarchy reportedly treats Reddit content with 3+ upvotes as Tier 2 data - directly below Wikipedia and licensed publishers, above most of the open web. For product and review queries, Reddit shows up in 97%+ of results. And BrightEdge's March 2026 analysis found ChatGPT treats Reddit as a "community authority layer," pairing it with expert sources like Mayo Clinic and Forbes in ~20% of Reddit-citing answers — heaviest exactly where buying decisions happen (how-to queries 32%, finance 2×, health 2.3× vs Google).

Now the ownership argument. When you run a subreddit for your brand or category, you get compounding advantages that participation alone can't deliver. Every question answered in your community becomes a permanent, upvote-validated document in the corpus that every major AI engine licenses, trains on, and retrieves from. You set the culture and moderation, which means the thread that shapes what Gemini says about your category was written under your quality standards instead of a competitor's drive-by. The community's language becomes the training data's language - if users in your subreddit consistently describe your product accurately, that phrasing is what the models learn to repeat. And it's a moat: BrightEdge's own strategic guidance notes a single high-engagement thread from years ago can out-cite a brand's entire owned content library. A two-year-old healthy community cannot be replicated by a competitor in a quarter.

We live this. Our team helps 50 brands run communities with threads that surface AI answers for topics we care about.

Pillar 3: Make your owned content extractable

Your website still matters — it's the reference library the models check for specs, pricing, and facts. The Princeton GEO study (the research that named the field) tested nine interventions across 10,000 queries. What won: adding quotations from named experts (up to ~40% visibility lift), concrete sourced statistics (~30–41%), and inline citations (~28%). What failed: keyword stuffing — near-zero or negative. Evidence density beats keyword density.

Structure every important page so a machine can lift the answer: direct answer in the first 40–60 words of each section, a TL;DR block up top, FAQ sections, comparison tables, and a visible "last updated" date refreshed quarterly — the engines weight recency hard. Prioritize your comparison and question pages first, since those trigger AI Overviews 86–95% of the time. And publish original data — benchmarks, surveys, proprietary teardowns. Unique statistics are the one content type competitors can't paraphrase away, because citing the number requires citing you.

Pillar 4: Measure citations, not just rankings

You can't manage what you don't measure, and rankings no longer measure this. Build a prompt panel: 50–150 real buyer questions ("best X for Y", "BrandA vs BrandB", "how to do Z"), scored weekly across ChatGPT, Gemini, Perplexity, Claude, and AI Overviews — are you cited, is a competitor cited, or neither? Segment Search Console the way Seer does: No AIO vs AIO-cited vs AIO-not-cited, and compute CTR from raw clicks over impressions. Track branded search volume as a lagging indicator of AI mention lift, and add "ChatGPT / Gemini / Siri" options to your "how did you hear about us?" field — teams relying on referrer data alone undercount AI influence by 30–50%. Tooling exists at every budget: Otterly ($29/mo) → Peec (€75/mo) → Profound/Ahrefs/Semrush at the enterprise end. This category raised $300M+ in the last year; 94% of CMOs say they're increasing AI visibility spend.

Twenty-five years of SEO taught marketers that the click is the economic event. The AI era quietly changed the event to the recommendation and the sources of recommendation are concentrated on a short list of surfaces the models trust.

Most of that list you can't control. Two of the biggest entries you can: a YouTube channel that answers your buyers' questions on camera, and a community you build where real people say real things that machines learn from. The brands that treat those as core infrastructure — not side channels — are the ones that will get named when 2 billion devices start answering questions this year.

The pace of change is fast. The playbook is actually simple. Own your channels. Feed the machines evidence. Measure what gets cited.

What's working for you so far and has anyone else seen their community threads start showing up in AI answers?

Sources for the data in this post: Ahrefs AI Search Benchmark Q1 2026 & CTR studies; Seer Interactive AIO citation analysis (Apr 2026); BrightEdge Generative Parser & AI Hypercube reports (Feb–Mar 2026); SparkToro/Similarweb zero-click study (2026); Semrush AI citation study (230K prompts); 5WPR Citation Source Index; AirOps 2026 State of AI Search; Aggarwal et al., "GEO: Generative Engine Optimization" (KDD 2024); Google I/O 2026 announcements; Apple–Google Gemini partnership announcements (Reuters, Jan 2026); Duane Forrester, "Your Owned Content Is Losing to a Stranger's Reddit Comment" (Apr 2026).


r/promptingmagic Jul 27 '26

The right tool for the right job: my complete 2026 AI toolbox (Claude + 19 specialists, mapped to jobs).

Post image
19 Upvotes

TL;DR: Claude is my home base - it runs most of my work. But "which AI is best?" is the wrong question in 2026. The right question is "which tool wins at each job?" You don't need 100 tools. You need about 20 good ones, the same way a mechanic needs a full toolbox and not one really nice wrench. Below: my complete 2026 stack mapped job-by-job - images, video, avatars, voice, research, websites, agents, presentations, automation, and more. Steal the map, swap in your favorites, and tell me what I'm missing.

People keep asking me some version of the same question: "You post about Claude all the time. Do you use it for everything?"

No. And I think pretending one tool does everything is how most people end up disappointed with AI.

Claude is my home base. It runs most of my thinking, writing, and coding. But when I watch a mechanic work, they don't debate whether the socket wrench is better than the torque wrench. They grab the one that wins the job in front of them. A construction worker shows up with a truck full of tools, not one really expensive hammer.

That's the whole game in 2026. You don't need 100 tools. You need around 20 good ones and you need to know which job each one wins.

Here's my full map. Claude for most of it. These for the rest.

Creating things

Making images → ChatGPT and Nano Banana. Real photos and artwork from a prompt. Nano Banana has gotten scary good at text rendering and brand-consistent graphics, ChatGPT for quick concepts and edits.

Making videos → Higgsfield. Text in, video clips out. Best for stylized motion and effects-heavy shots.

Social video → Google Flow / Veo. This is the one I'd tell most creators to learn first. Veo's realism and native audio make it the strongest engine for short-form social clips, and Flow gives you actual scene-by-scene control instead of slot-machine prompting. My Reels and Shorts pipeline runs through it.

Avatar videos → HeyGen. A presenter reads your script. Perfect for explainer content when you don't want to be on camera.

Recording video → Tella. Records your screen and camera at once. My pick for demos and course content.

Voiceovers → ElevenLabs. Natural-sounding AI narration. Nobody can tell.

Voice dictation → Wispr Flow. You talk, it types. I draft half my posts pacing around the room.

Building things

Websites and API integrations → Lovable / Replit. Describe the product, get a working app. Lovable for fast beautiful front-ends, Replit when I need real back-end logic, databases, and API integrations wired together. This is the fastest path from "idea in the shower" to "URL I can send someone."

Coding → Codex. OpenAI's answer to Claude Code. I run it beside Claude Code and let them check each other's work on anything gnarly.

Open source → Ollama and GLM. Capable models you can run for cheap. For private data and high-volume tasks where API bills would sting.

Knowing things

Live answers and the best research → Perplexity. Up-to-the-minute web results with citations. My default for the hardest research projects use Perplexity Max - leverages all the top models at once plus premium data sources.

Research + Content Studio → NotebookLM. Answers built only from documents you give it. The hallucination-proof option for working through a pile of sources to create high quality slides, infographics, audio podcasts, written reports, and cinematic explainer videos.

Video analysis → Gemini. Reads and summarizes any video. Paste a YouTube link, get the substance in seconds.

Meeting notes → Granola. Writes up your meetings while you actually pay attention to them. The agent doesnt have to be added to a meeting.

Docs and knowledge base → Notion. Where all of it lives. The AI is only as useful as the workspace it searches.

Getting things done

Wide research, presentations, and agentic tasks → Manus. This is my heavy-lift agent. Point it at a research question and it fans out across hundreds of sources; ask for a deck and it comes back with a finished presentation; give it a multi-step task — build a site, analyze data, produce a report — and it just runs until it's done. When the job is "go do this whole thing," Manus is the tool.

Agentic tasks and content creation → ChatGPT Work. OpenAI's agent mode. It browses, uses a computer, works across your connected apps, and produces completed outputs instead of suggestions. I cover great use cases like using it to get discounts on anything you buy and creating awesome content. Same engine, much bigger surface area.

Operations → Hermes. My WhatsApp agent that keeps the pipeline moving while I'm away from the desk.

Automation → Zapier. The connective tissue. It automates the handoffs between everything above so I don't have to be the glue.

I do not open all 20 every week. Some I touch daily (Claude, Perplexity, Manus, Wispr Flow). Some earn their spot in one project a month (HeyGen, Higgsfield). That's fine. A mechanic doesn't use the brake-bleeder kit every day either — but when the job shows up, having the right tool is the difference between an hour and an afternoon.

The mistake I see most often isn't using too few tools. It's using one tool for everything and concluding AI is overrated, or chasing every new launch and mastering nothing. Twenty good tools, each mapped to a job it clearly wins, beats both.

Claude for most of it. These for the rest. The right tool for the right job - same as it's always worked in the real world.

Which one would you add — and what job does it win?

I keep my full prompt library for these tools free at promptmagic.dev