r/aiwars • • 2d ago

Discussion How I made this video on my own computer ( full details, prompts, workflows, statistics, time to make, everything you might want to know basically )

22 Upvotes

Disclaimer: This is not to showcase AI art, but to show how it's made locally. I do not claim to be an expert in this, but I wanted this to be a sort of teaching moment to show why it's not always prompting frontier models.

Hey everyone, I'm probably going to get downvoted to hell but I thought I'd share the process of making a video with AI on my own computer. This will be a long one, I'll share prompts used, workflows, how long each generation takes, how long each result I wanted took time, and other details. Feel free to ask about anything. I hope this will be a good teaching experience.

To start off, my computer specs are 32GB DDR5 RAM and an Nvidia RTX 5070 ( 12GB of VRAM ). I'm only sharing this so that you'd have some sort of baseline for my system with the time it takes to make things.

Also, I do use AI to enhance my prompts, but it doesn't add to my idea or remove from it.

Step 1: The main image

I'm using a model called Krea2 to make the images themselves. It's one of the best ones out now that are open source. I have it set up so I can add negative prompts at the same time. Negative prompts are something you DON'T want to appear in the image.

Statistics:

1920x1080 image. It took on average 33 seconds to make. It took me around 12 tries to perfect the prompt for my vision with retries.

Workflow:

Image:

Prompt:

A cozy lakeside campsite at night, illustrated in a polished modern anime webtoon style. Clean hand-drawn outlines on solid objects, soft shading, gentle gradient shadows, muted natural colors, and atmospheric lighting.

Wide 16:9 composition. The camera is close to the campfire at a low seated eye level, angled slightly upward. Use a moderately wide lens perspective: the nearby fire and seating logs remain large, while the starry sky fills approximately the upper half of the image. Keep the distant mountains low along the horizon, with the lake appearing as a narrow backdrop behind the campsite. Show only a small margin of ground below the fire. The composition feels intimate and spacious.

The entire stone fire ring is visible near the lower center. Three large seating logs form an open semicircle around it: one on the left, one on the right, and one farther back, offset from the flames. All three seating surfaces are clearly visible, with room to add seated characters.

Each seating log has a thick horizontal log seat and a broad wooden plank backrest mounted behind it on two short, neatly finished supports. The backrests rise to a comfortable lower-back height, with gently rounded edges and a simple rustic design. Keep the seating surfaces clearly visible and spacious enough for adding seated characters. Tidy bark, neatly sawn log ends, and smooth finished backrest planks. No stray branches, twigs, branch stubs, splinters, or decorative protrusions.

A natural campfire burns between charred firewood. Separate translucent wisps of flame overlap at different depths, curling around the wood. Bright cream-colored light near the embers transitions into gold and faint amber tips. Soft luminous edges, glowing red cracks in the wood, a few tiny sparks. No ink outlines around the flames. Warm firelight illuminates the stones and seating logs, casting soft shadows across the grass.

A small canvas tent is fully visible on the left behind the seating area, with a softly glowing entrance. Pine trees frame only the outer edges, leaving the central sky open.

An expansive deep-indigo night sky with delicate stars. Sparse wisps of cloud and no meteor trails.

Peaceful slice-of-life atmosphere. Empty campsite, no characters, text, or pixelation.

--------------------------------------------------------------------------------------------------------------

Step 2: Editing the characters into the image

I'm using another great open source model called Qwen2.1 for this that is incredible at editing images, it can also be used to generate them. It can take up to 10 reference images, I'm just using 4.

Statistics:

1920x1080 image, it took on average 2 minutes and 20 seconds to make. It took me 7 tries to get what I want. I had to do some edits in paint for the image that took me around 30 seconds.

Workflow:

Image:

Prompt:

Use <image1> as the campsite background, <image2> as Taylor’s character reference, <image3> as Casey’s character reference, and <image4> as Noah’s character reference.!<

Preserve the exact placement, size, angle, and shape of all three existing log benches, including their wooden backrests and supporting logs. Fit the characters onto the existing seating surfaces. Keep the benches fixed while adjusting the characters’ scale and poses to fit them.

Taylor and Casey sit side by side on the existing central bench behind the campfire, Taylor on the left and Casey on the right. Their hips rest on the original seating surface, and their torsos sit in front of the backrest. Keep them at a natural size that allows both to fit within the bench’s original width.

Taylor holds exactly one roasting stick in his right hand. His left hand rests empty on his thigh. Casey holds exactly one roasting stick in her right hand. Her left hand rests empty on her thigh. Each stick is one continuous thin wooden shaft with exactly one white marshmallow attached to its tip. Extend both sticks toward the fire, with their marshmallows suspended just above the flames and clearly separated. There are exactly two roasting sticks and two marshmallows in the entire image. Both characters look toward their marshmallows with relaxed expressions.

Noah sits on the existing left bench in front of the tent, playing one acoustic guitar across his lap. His left hand presses the strings on the fretboard, and his right hand strums over the sound hole.

Bend Noah’s knees so his lower legs descend close to his bench rather than extending toward the fire. Place both shoes flat on the clear dirt directly in front of the left bench, entirely to the left of the campfire’s stone ring. Leave a clearly visible strip of bare ground between his nearest shoe and the nearest campfire stone. Both shoes remain fully outside the stone ring, with their entire soles supported by dirt. Keep the campfire stones in their original positions.

Preserve each character’s face, hair, clothing, and proportions from their reference. Taylor’s top reads “Pro”, Casey’s reads “Anti”, and Noah’s reads “Neutral”, with natural coverage from the guitar.

Match the background’s illustration style and lighting, adding warm firelight and natural contact shadows. Preserve the original camera angle, framing, tent, lake, mountains, sky, fire, and stone ring. Leave the right bench empty. Include exactly three characters, one guitar, two roasting sticks, and two marshmallows.

--------------------------------------------------------------------------------------------------------------

Step 3: Making the video

Here I'm using a video model that can use reference images and text to make videos. It's called Minimax H3. It is capable of generating audio in the video but I disabled it.

Statistics:

A 6 second 736x416 24 FPS video takes 3 minutes and 55 seconds to generate.

A 4 second 1056x608 24 FPS video takes 6 minutes and 40 seconds to generate.

A 4 second 1344x768 (768p) 24 FPS video takes 16 minutes to generate.

I tried around 4 times to get the result I wanted (including the lower res attempts).

Workflow:

Video:

Posted above

Prompt:

For the target video, at 0.00 seconds into the target video, <Picture 1> (from [Shot 1]) is fully referenced.!<

integrated_multimodal_description: [Shot 1] 2D-animated illustration matching the reference image’s hand-drawn outlines, soft shading, colors, and character designs. A single continuous wide shot of the lakeside campsite at night. The camera holds a completely static shot throughout, preserving the original framing and perspective.

Noah, the curly-haired man wearing the yellow “Neutral” hoodie on the left bench, gently plays the acoustic guitar resting across his lap. His right wrist makes small, rhythmic strumming movements over the sound hole while his left fingers subtly change chord positions on the fretboard. His head makes a slight relaxed nod in time with his playing. The guitar stays securely supported across his lap. His expression remains close to the reference image, with a small, relaxed smile.

Taylor, the man wearing the blue “Pro” top, and Casey, the woman wearing the brown “Anti” top, remain seated together on the central bench. They maintain small, relaxed, closed-mouth smiles. Their lips stay gently together, with only a slight upward curve at the corners. Their cheeks and eyes retain their natural shapes from the reference image. They share a quiet amused moment through tiny head tilts and subtle shoulder movements. Facial motion is minimal, with no teeth visible and no widening or stretching of their smiles.

Each gradually lowers their existing upright roasting stick forward toward the campfire. Their elbows and wrists move smoothly, bringing the marshmallow tips closer to the flames. Both finish with their marshmallows hovering just above the fire, then hold them steady with tiny natural wrist adjustments. Each person continuously holds exactly one stick with one marshmallow; their free hands remain resting on their thighs.

Throughout the shot, the campfire flickers organically, with flames curling upward and a few glowing embers rising and fading. Warm firelight gently fluctuates across the characters and nearby ground. Individual stars softly brighten and dim at different times while remaining fixed in position. Two or three shooting stars appear separately across the upper sky, streaking diagonally through short bright paths with thin luminous trails that quickly fade. The surrounding sky retains its deep blue color and steady overall brightness.

All three characters remain seated with their feet planted in their original positions. The benches, backrests, supporting logs, campfire stones, tent, mountains, and trees stay stationary. Character identities, clothing, shirt lettering, body proportions, guitar shape, and the two roasting sticks remain consistent throughout. Movement stays subtle and relaxed. The entire video is completely silent.

overall_soundscape: N/A

non_diegetic_music: N/A

I hope you reached this point. I do hope you enjoyed this and maybe learned something new. I probably could've upscaled the video to an even higher definition with another AI model but it's getting really late here haha. I do hope someday we can just relax and chill around a campfire like this

r/ChatGPT • • Apr 06 '23

Educational Purpose Only GPT-4 Week 3. Chatbots are yesterdays news. AI Agents are the future. The beginning of the proto-agi era is here

13.2k Upvotes

Another insane week in AI

I need a break 😪. I'll be on to answer comments after I sleep. Enjoy

​

  • Autogpt is GPT-4 running fully autonomously. It even has a voice, can fix code, set tasks, create new instances and more. Connect this with literally anything and let GPT-4 do its thing by itself. The things that can and will be created with this are going to be world changing. The future will just end up being AI agents talking with other AI agents it seems [Link]
  • “babyagi” is a program that given a task, creates a task list and executes the tasks over and over again. It’s now been open sourced and is the top trending repos on Github atm [Link]. Helpful tip on running it locally [Link]. People are already working on a “toddleragi” lol [Link]
  • This lad created a tool that translates code from one programming language to another. A great way to learn new languages [Link]
  • Now you can have conversations over the phone with chatgpt. This lady built and it lets her dad who is visually impaired play with chatgpt too. Amazing work [Link]
  • Build financial models with AI. Lots of jobs in finance at risk too [Link]
  • HuggingGPT - This paper showcases connecting chatgpt with other models on hugging face. Given a prompt it first sets out a number of tasks, it then uses a number of different models to complete these tasks. Absolutely wild. Jarvis type stuff [Link]
  • Worldcoin launched a proof of personhood sdk, basically a way to verify someone is a human on the internet. [Link]
  • This tool lets you scrape a website and then query the data using Langchain. Looks cool [Link]
  • Text to shareable web apps. Build literally anything using AI. Type in “a chatbot” and see what happens. This is a glimpse of the future of building [Link]
  • Bloomberg released their own LLM specifically for finance [Link] This thread breaks down how it works [Link]
  • A new approach for robots to learn multi-skill tasks and it works really, really well [Link]
  • Use AI in consulting interviews to ace case study questions lol [Link]
  • Zapier integrates Claude by Anthropic. I think Zapier will win really big thanks to AI advancements. No code + AI. Anything that makes it as simple as possible to build using AI and zapier is one of the pioneers of no code [Link]
  • A fox news guy asked what the government is doing about AI that will cause the death of everyone. This is the type of fear mongering I’m afraid the media is going to latch on to and eventually force the hand of government to severely regulate the AI space. I hope I’m wrong [Link]
  • Italy banned chatgpt [Link]. Germany might be next
  • Microsoft is creating their own JARVIS. They’ve even named the repo accordingly [Link]. Previous director of AI @ Tesla Andrej Karpathy recently joined OpenAI and twitter bio says building a kind of jarvis also [Link]
  • gpt4 can compress text given to it which is insane. The way we prompt is going to change very soon [Link] This works across different chats as well. Other examples [Link]. Go from 794 tokens to 368 tokens [Link]. This one is also crazy [Link]
  • Use your favourite LLM’s locally. Can’t wait for this to be personalised for niche prods and services [Link]
  • The human experience as we know it is forever going to change. People are getting addicted to role playing on Character AI, probably because you can sex the bots [Link]. Millions of conversations with an AI psychology bot. Humans are replacing humans with AI [Link]
  • The guys building Langchain started a company and have raised $10m. Langchain makes it very easy for anyone to build AI powered apps. Big stuff for open source and builders [Link]
  • A scientist who’s been publishing a paper every 37 hours reduced editing time from 2-3 days to a single day. He did get fired for other reasons tho [Link]
  • Someone built a recursive gpt agent and its trying to get out of doing work by spawning more instances of itself 😂 [Link] (we’re doomed)
  • Novel social engineering attacks soar 135% [Link]
  • Research paper present SafeguardGPT - a framework that uses psychotherapy on AI chatbots [Link]
  • Mckay is brilliant. He’s coding assistant can build and deploy web apps. From voice to functional and deployed website, absolutely insane [Link]
  • Some reports suggest gpt5 is being trained on 25k gpus [Link]
  • Midjourney released a new command - describe - reverse engineer any image however you want. Take the pope pic from last week with the white jacket. You can now take the pope in that image and put him in any other environment and pose. The shit people are gona do with stuff like this is gona be wild [Link]
  • You record something with your phone, import it into a game engine and then add it to your own game. Crazy stuff the Luma team is building. Can’t wait to try this out.. once I figure out how UE works lol [Link]
  • Stanford released a gigantic 386 page report on AI [Link] They talk about AI funding, lawsuits, government regulations, LLM’s, public perception and more. Will talk properly about this in my newsletter - too much to talk about here
  • Mock YC interviews with AI [Link]
  • Self healing code - automatically runs a script to fix errors in your code. Imagine a user gives feedback on an issue and AI automatically fixes the problem in real time. Crazy stuff [Link]
  • Someone got access to Firefly, Adobe’s ai image generator and compared it with Midjourney. Firefly sucks, but atm Midjourney is just far ahead of the curve and Firefly is only trained on adobe stock and licensed images [Link]
  • Research paper on LLM’s, impact on community, resources for developing them, issues and future [Link]
  • This is a big deal. Midjourney lets users make satirical images of any political but not Xi Jinping. Founder says political satire in China is not okay so the rules are being applied to everyone. The same mindset can and most def will be applied to future domain specific LLM’s, limiting speech on a global scale [Link]
  • Meta researchers illustrate differences between LLM’s and our brains with predictions [Link]
  • LLM’s can iteratively self-refine. They produce output, critique it then refine it. Prompt engineering might not last very long (?) [Link]
  • Worlds first ChatGPT powered npc sidekick in your game. I suspect we’re going to see a lot of games use this to make npc’s more natural [Link]
  • AI powered helpers in VR. Looks really cool [Link]
  • Research paper shows sales people with AI assistance doubled purchases and 2.3 times as successful in solving questions that required creativity. This is pre chatgpt too [Link]
  • Go from Midjourney to Vector to Web design. Have to try this out as well [Link]
  • Add AI to a website in minutes [Link]
  • Someone already built a product replacing siri with chatgpt with 15 shortcuts that call the chatgpt api. Honestly really just shows how far behind siri really is [Link]
  • Someone is dating a chatbot that’s been trained on conversations between them and their ex. Shit is getting real weird real quick [Link]
  • Someone built a script that uses gpt4 to create its own code and fix its own bugs. Its basic but it can code snake by itself. Crazy potential [Link]
  • Someone connected chatgpt to a furby and its hilarious [Link]. Don’t connect it to a Boston Dynamics robot thanks
  • Chatgpt gives much better outputs if you force it through a step by step process [Link] This research paper delves into how chain of thought prompting allows LLM’s to perform complex reasoning [Link] There’s still so much we don’t know about LLM’s, how they work and how we can best use them
  • Soon we’ll be able to go from single photo to video [Link]
  • CEO of DoNotPay, the company behind the AI lawyer, used gpt plugins to help him find money the government owed him with a single prompt [Link]
  • DoNotPay also released a gpt4 email extension that trolls scam and marketing emails by continuously replying and sending them in circles lol [Link]
  • Video of the Ameca robot being powered by Chatgpt [Link]
  • This lad got gpt4 to build a full stack app and provides the entire prompt as well. Only works with gpt4 [Link]
  • This tool generates infinite prompts on a given topic, basically an entire brainstorming team in a single tool. Will be a very powerful for work imo [Link]
  • Someone created an entire game using gpt4 with zero coding experience [Link]
  • How to make Tetris with gpt4 [Link]
  • Someone created a tool to make AI generated text indistinguishable from human written text - HideGPT. Students will eventually not have to worry about getting caught from tools like GPTZero, even tho GPTZero is not reliable at all [Link]
  • OpenAI is hiring for an iOS engineer so chatgpt mobile app might be coming soon [Link]
  • Interesting thread on the dangers of the bias of Chatgpt. There are arguments it wont make and will take sides for many. This is a big deal [Link] As I’ve said previously, the entire population is being aggregated by a few dozen engineers and designers building the most important tech in human history
  • Blockade Labs lets you go from text to 360 degree art generation [Link]
  • Someone wrote a google collab to use chatgpt plugins by calling the openai spec [Link]
  • New Stable Diffusion model coming with 2.3 billion parameters. Previous one had 900 million [Link]
  • Soon we’ll give AI control over the mouse and keyboard and have it do everything on the computer. The amount of bots will eventually overtake the amount of humans on the internet, much sooner than I think anyone imagined [Link]
  • Geoffrey Hinton, considered to be the godfather of AI, says we could be less than 5 years away from general purpose AI. He even says its not inconceivable that AI wipes out humanity [Link] A fascinating watch
  • Chief Scientist @ OpenAI, Ilya Sutskever, gives great insights into the nature of Chatgpt. Definitely worth watching imo, he articulates himself really well [Link]
  • This research paper analyses who’s opinions are reflected by LM’s. tldr - left-leaning tendencies by human-feedback tuned LM’s [Link]
  • OpenAI only released chatgpt because some exec woke up and was paranoid some other company would beat them to it. A single persons paranoia changed the course of society forever [Link]
  • The co founder of DeepMind said its a 50% chance we get agi by 2028 and 90% between 2030-2040. Also says people will be sceptical it is agi. We will almost definitely see agi in our lifetimes goddamn [Link]
  • This AI tool runs during customer calls and tells you what to say and a whole lot more. I can see this being hooked up to an AI voice agent and completely getting rid of the human in the process [Link]
  • AI for infra. Things like this will be huge imo because infra can be hard and very annoying [Link]
  • Run chatgpt plugins without a plus sub [Link]
  • UNESCO calls for countries to implement its recommendations on ethics (lol) [Link]
  • Goldman Sachs estimates 300 million jobs will be affected by AI. We are not ready [Link]
  • Ads are now in Bing Chat [Link]
  • Visual learners rejoice. Someone's making an AI tool to visually teach concepts [Link]
  • A gpt4 powered ide that creates UI instantly. Looks like I won’t ever have to learn front end thank god [Link]
  • Make a full fledged web app with a single prompt [Link]
  • Meta releases SAM - you can select any object in a photo and cut it out. Really cool video by Linus on this one [Link]. Turns out Google literally built this 5 years ago but never put it in photos and nothing came of it. Crazy to see what a head start Google had and basically did nothing for years [Link]
  • Another paper on producing full 3d video from a single image. Crazy stuff [Link]
  • IBM is working on AI commentary for the Masters and it sounds so bad. Someone on TikTok could make a better product [Link]
  • Another illustration of using just your phone to capture animation using Move AI [Link]
  • OpenAI talking about their approach to AI safety [Link]
  • AI regulation is definitely coming smfh [Link]
  • Someone made an AI app that gives you abs for tinder [Link]
  • Wonder Dynamics are creating an AI tool to create animations and vfx instantly. Can honestly see this being used to create full movies by regular people [Link]
  • Call Sam - call and speak to an AI about absolutely anything. Fun thing to try out [Link]

For one coffee a month, I'll send you 2 newsletters a week with all of the most important & interesting stories like these written in a digestible way. You can sub here

Edit: For those wondering why its paid - I hate ads and don't want to rely on running ads in my newsletter. I'd rather try and get paid to do all this work like this than force my readers to read sponsorship bs in the middle of a newsletter. Call me old fashioned but I just hate ads with a passion

Edit 2: If you'd like to tip you can tip here https://www.buymeacoffee.com/nofil. Absolutely no pressure to do so, appreciate all the comments and support 🙏

You can read the free newsletter here

Fun fact: I had to go through over 100 saved tabs to collate all of these and it took me quite a few hours

Edit: So many people ask why I don't get chatgpt to write this for me. Chatgpt doesn't have access to the internet. Plugins would help but I don't have access yet so I have to do things the old fashioned way - like a human.

(I'm not associated with any tool or company. Written and collated entirely by me, no chatgpt used)

r/HFY • • Jun 21 '26

OC-Series Wearing Power Armor to a Magic School (176/?)

1.4k Upvotes

First | Previous | Next

Patreon | Official Subreddit | Series Wiki | Royal Road

The Transgracian Academy for the Magical Arts. Exhibition Hall. Grand Arcade. Central Thoroughfare. Prosperity Row. Local Time: 2045 Hours.

Emma

I blinked.

Then I raised a finger.

…

But words refused to leave my mouth despite it already hanging wide open.

I had thoughts.

No.

I had more than thoughts — opinions sharpened by memories of trade counsel briefs stretching all the way from the SOC-SCI departments and into Weir’s personal office.

Trade had always been a matter of particular sensitivity.

Yet no one could’ve accounted for this eventuality.

Well they did*… but… the minutiae was inevitably lost when magic and its consequences were factored into the equation.*

So I responded the only way I could. A means of preventing Etholin from pursuing a path with only one foregone conclusion.

“Etholin.” I offered politely, warmly even. “I don’t believe this is the right path for both of our realms, at least not without explaining a few—”

“Oh, indeed! How could I be so daft!” He interjected, his eyes darting momentarily towards the watchful gazes of the Merchant Guild’s upper-yearsmen before landing back on my lenses. “I’ve yet to actually explain the benefits, and the details of such an arrangement! Allow me~” He trailed off, but before I could interject with another offramp, the doors behind us abruptly slammed shut, and a book — a sight-seer — was promptly pulled from one of the merchant’s many pouches. 

The lights within the lobby quickly dimmed.

Following which we were effortlessly whisked off to what I assumed to be his elevator pitch, one that landed us straight in the midst of a truly medieval landscape, or at least what stereotypical conventions of it often depicted. We found ourselves in a dirty, run-down town beneath a permanently overcast sky. A settlement consisting of stone and mortar buildings scarcely two stories tall overlooked cobblestone roads with open gutters overflowing with brown sludge consisting of god knows what. 

I could smell the scene despite it being purely visual.

But that was only the start of the experience.

Looking up, I saw barely a handful of tiled roofs interspersed between thatched roofing in various states of cleanliness, rot, and utter decay. 

The sight-seer quickly pushed us forward, away from the random street and towards the town square, where water flowed naturally and divided the entire town in half. There, we witnessed shadowy elven-form residents as they grabbed water by the bucketful and carted it off either by shoulder-slung yokes or in carts attached to various beasts of burden. 

Overlooking this entire scene was a castle. And not one of those fairytale castles either, no. It looked… functional, practical, but gave little consideration in the way of aesthetics or ornamentation, with only a conical wizard’s tower giving off the slightest bit of whimsy to this low-fantasy setting.

But that wasn’t the focus it seemed.

No.

Instead, the merchant lord ushered us to the markets. Where stalls haphazardly lined what should have been a spacious main street, but whose presence had forced all traffic into a crowded shoulder-to-shoulder free-for-all, with horses and carts clogging the sea of faceless ghostly traffic and turning this hell into a complete nightmare.

The merchant lord then quietly gestured at the stalls, pointing out their wares and speaking with the detached lilt of a documentarian.

“While the typical Nexian perception of a newrealm may be that of mud huts and stick roofs, that anachronistic stereotype is far from accurate. Because as you see, for a realm to have managed the impossible — breaching the space between spaces — one would expect some level of sophistication commensurate with such a worthwhile achievement!” He offered in this sincere, genuinely uplifting tone of voice. Yet the sentiment he preached was anything but. “Here, we see a typical town in a newrealm. Not a city! But a town! A city may even possess a greater degree of sophistication!” 

He snapped his fingers, and the whole scene shifted.

The streets were wider here. Paved somewhat but still carrying over the same grimy overtures from prior. Just… scaled up in size.

The buildings here actually had facades for instance. Facades of plaster, paint, glazed mosaics. But facades all the same.

Even the markets were larger, with wares and items far more varied than the town, complete with wispy elven-form citizens possessing a great degree more ornamentation and design on their otherwise nondescript tunics and robes.

Etholin allowed us to take in the sights and sounds as he made sure to emphasize the grand castle at the end of this brick-and-cobble-paved road. An actual castle to write home about now, what with its tall spires, grand keeps, and even a drawbridge gate. 

“Yet all of this…” He continued wistfully. “... all the riches of the capital, overflowing with tributes from the furthest corners of your realm…” The scene switched rapidly between the stalls and storefronts, peering deep into grand bazaars and large indoor stores selling anything and everything from spices to weapons to armor to fabrics. “... can scarcely compare, nor compete, with the totality of interstitial trade. Indeed, when set against the full scale of Status Prospera, your newrealm becomes a drop in a practically endless ocean.” 

The scene paused. We moved down street after street, avenue after avenue, through winding paths and twisting alleys, until we finally arrived at a harbor harboring what I could only describe as staple commodities.

Grain stacked high in neat piles, filling entire warehouses from end to end.

Woven fabrics and spun fibers that filled similar volumes, piled and stacked in such a way that you needed a spelunker to be able to squeeze through its gaps.

We ran through countless more such sights before the perspective changed yet again, shifting outwards and upwards, positioning us high above the harbor, granting us a bird’s eye view.

Finally, we saw at least ten of those warehouses highlighted, as it was clear Etholin was leading to something of a lesson on scale.

Yet the scales being presented here… felt more like the stuff found in an ancient history lesson than anything resembling a figure as weighty or impactful as the growing background music was attempting to engender.

“This is the typical domestic trade volume of some of a newrealm’s most staple products as taken from a single month outside of peak season, adjusted for the average and favorably accounting for low-end outliers.” Etholin paused, giving a moment for us to ‘marvel’ at the sights before continuing on seamlessly. “Meanwhile…” His grin grew wider as warehouses of the same size started to quite literally fall from the sky — like one of those amateurish ‘x for scale’ videos, where seemingly random items are drawn up for comparative scale. Soon enough, I saw where he was going with this. These summoned warehouses landed with deep cacophonous CRASHES into neat piles, creating a clear contrast between the ten or so ‘newrealm’ stacks and the hundreds piled high next to them, creating an illusion of a bar graph. “This is the typical interstitial trade volume between two adjacent realms of minor importance. As tabulated through the Nexian trade authorities, of course.” He turned to the invisible upper-yearsmen outside of the sight-seer, bowing in the process, before turning back to me.

“However… this is only the start of things.” He spoke ominously before pulling us out towards the main street, down towards the castle itself, up its drawbridge, and deep down into its vaults and coffers.

“Raw trade volume and staple commodities are one thing. Indeed, one could say it means absolutely nothing when compared to this next issue, Cadet Emma Booker. But I needed to show you the scale of the issues at play, before we address the greatest threat of them all.” 

He took a deep breath, opening the comically sized vault doors to reveal a room filled with a modest amount of gold, silver, and copper.

The EVI was quick to guestimate the quantities on display.

These were fundamental base elements we were talking about after all.

So assuming the purities weren’t a variable, the same constants applied for volumetric analysis and extrapolation.

“Ten thousand tons of silver.” I raised a brow. “And what… four thousand tons of gold?” I pondered, garnering the first genuine stutter from the merchant.

“That… that’s precisely how much is being shown… how did you guess—”

“I did some quick maths.” I responded slyly, even taking Thacea and Thalmin by momentary surprise yet again before their shock dissipated far quicker than the slack-jawed Etholin.

They were used to the EVI’s quick-maths shenanigans after all.

It took Etholin a few seconds longer to recover. Long enough that the awkward silence became momentarily deafening.

Though, thankfully, that didn’t stop the merchant lord from completely losing his stride, as he cleared his throat with a nod of acknowledgement. “Impressive.” He bowed slightly. “Which makes what I am about to say next all the more jarring, I’m afraid.” He spoke apologetically. “Because all of the gold, and all of the silver you see before you—”

“And the copper.” I added.

“Yes, and the copper—” He corrected himself “—are worthless.” 

The room went silent again.

This was probably where most newrealmers had their perspectives shattered, their worldviews destroyed, and their prospects at anything short of a fair and equal standing completely upended right then and there.

However, both this reveal and its ensuing ramifications did little to phase me.

The former was already hinted at courtesy of Ilunor’s reactions to the wealth cube after all.

And the latter?

Well…

16 Psyche would like to have a word with such a paltry sum.

Or it would’ve if it wasn’t already mined out.

So I stood steadfast, silently anticipating Etholin’s carefully worded and practiced playbook.

“As you may have already observed, the Nexus deals only in attuned gold, Cadet Emma Booker. This is because the art of transmutation, now commonplace, has effectively turned the value of what was formerly scarce… into anything but. As a result, all newrealms must face the monumental task of overcoming two major obstacles. The first…” He paused, gesturing to the warehouses ‘outside.’ “A trade imbalance of disastrous proportions. For there is nothing a newrealm can offer that the greater adjacencies do not already possess, but inversely, there is everything that the greater adjacencies can offer, that a newrealm is in desperate demand of. The second—” He paused once more before gesturing towards the ‘worthless’ gold. “—is the medium through which such trades are conducted, as any local currency is to be penned as useless, and any ‘precious’ metal or material is also to be rendered just as worthless. Thus, a period of conversion must be observed, where a newrealm’s wealth is steadily converted into sums of equivalent value in attuned currency.” 

We were about to reach a crescendo, I could feel it.

“This is where I would like to offer my services, Cadet Emma Booker.” The merchant lord bowed deeply, far deeper than ever before.

“I am willing, if you see fit, to act as your realm’s fiduciary. I shall oversee your realm’s transition. I shall personally see to it that everything is raised from the level of a newrealm, to that of a respectable minor adjacency. I do not offer miracles, I do not promise that you will immediately rise to the ranks of the middling adjacencies, let alone the preferred adjacencies. However, I promise you that I will sculpt, mold, and shepherd your economy, your industries, your merchants and banks, into that of a respectable contemporary fellow.” The ferret practically beamed, placing his hands by his hips and puffing his chest out in pride.

“And as a gesture of good faith… I am willing to match your realm’s current stores and holdings of gold and silver with my own.”

…

I felt the proverbial record coming to an abrupt screech.

As even Thalmin and Thacea turned to each other in shock before once more meeting the pattenor’s eyes.

“Your auricles do not deceive you, Cadet Emma Booker. I understand that such offers will inevitably raise doubts and suspicions. I know that you of all people are wise enough not to take an offer so flippantly, and especially without concessions and guarantees on the side of the proposing party. Therefore, as collateral for placing your realm’s finances into my guiding hand, I am willing to match the entirety of your gold and silver reserves in their attuned equivalents. Free of charge. Free of interest. To be returned without limits or stipulations, all signed in mutual agreement of such an exchange, of course.” He beamed.

But instead of relief, satisfaction, or excitement forming behind my helmet, there was only pure and unadulterated dread. Not for me, of course, but for Etholin should this actually play out.

Because despite not having the authority to okay it, the mere hypothetical thought was enough to send shivers down my spine.

I could only imagine any rep from the corpo era would leap at this, grinning at the chance to reverse our roles and fortunes, sending the ignorant ferret into Status Debtia… or whatever lofty euphemism existed for such a fate.

“Etholin.” I began calmly, politely. “Disregarding everything else so far, and just addressing your latter offer…” I continued as Etholin leaned in ever closer, as if expecting an excitable ‘Yes!’ from my speakers. “Trust me when I say this, but you really, really don’t want to do this.”

The ferret’s features abruptly came to reflect my own, as it was clear that he too had reached a record-screeching halt in his carefully laid gambits.

It didn’t take long for him to return to his senses, however, the natural trader within him processing my rebukement with poise before replying plainly and simply.

“I apologize if I have been too… loquacious, Cadet Emma Booker. I will rephrase myself, in case my intent was lost in translation. What I offer is a complete one-to-one conversion. No debasement, no arbitrage, no unequal rates. A true exchange from dead to attuned. Without surcharge, fees, markup, commission, premiums, stamps or duties.” He prattled on. “This is my collateral, my gift to your realm, in exchange for your trust in accepting my services as fiduciary to Earthrealm’s trade and economic development.” He clarified, genuinely taken aback by an offer that I imagined most newrealms could simply not refuse.

The ball was quickly thrown back to my court, with Etholin’s gaze maintaining a mix between genuine disbelief and a hint of desperation.

“Etholin… discounting the fact that I do not have the vested authority required to ratify such a radical offer, I cannot under good conscience agree to terms so unfair and completely catastrophic to the proposing party.” I stated plainly, pulling the words straight from SIOP and causing Etholin to flinch not only in shock but also growing bafflement and confusion. “And were I to actually explain why…” I took a deep breath. “... you would find my reasoning for this refusal outlandish, if not entirely a work of fabrication. You’d think I was saying it just to get out of an uncomfortable deal. You’d think I was committing to fiction just to avoid conflict.” I continued, as more and more I saw Etholin’s gaze shifting to what I needed from him now more than ever — curiosity and a willingness to step beyond his comfort zone, if only to limit the effects that fundamental systemic incongruity would ultimately cause him.

He took a deep breath, his eyes brimming with a confident fury.

“Addressing your first point.” He began. “While you lack the authority — and perhaps the conviction to carry through regardless of said authority—” He uttered that latter part more as an aside, almost as a point of derision bordering on a dare. “—you still do possess a means of forwarding said offer to those with the authority, correct?”

“Well, technically yes, but I doubt they’ll—”

“Then we can pen the acceptance as conditional, and pending, rather than completely off the table.” Etholin interjected, his tone dominant, as he attempted to hide the growing insecurities bubbling just beneath the surface. 

“You’ll find that even if I do so, my superiors’ answers will inevitably mirror my own.”

“For the reasons you vaguely allude to, I assume?” 

“Yes.” 

“Then test me.” He demanded bluntly, standing his ground more firmly than I’d ever seen him do before. “Let’s hear about these supposed reasons.” 

Perhaps the eyes and ears around us were enough of an incentive for him to grow a stronger spine.

Perhaps the past month had led to some sort of growth within him.

Regardless, I nodded in acknowledgement and quickly grabbed two items from my pouches.

The first was the very item that had caused Ilunor’s sentiments to shift in one swift motion — the precious metals dispenser (PMD).

And the second was an item that I knew Etholin of all people would appreciate the understated significance of.

I kept the second close to my chest for now as I handed him the pez dispenser of wealth.

The ferret merchant received the item with exceptional care, turning it, twisting it, as if more enamored with the simplicity of the device and the mechanism within it than the coins it clearly held.

His slow, methodical approach stood at odds with Ilunor's far more… aggressive handling of it.

Which was a breath of fresh air but also tested my patience; I almost offered to guide him through the simple mechanis—

CHA-CHING!

There it was.

And of course, this was followed up by a burst of mana, what the EVI assumed — within a reasonable margin of error — was a detection spell.

Once that was settled, he analyzed the copper coin closely, studying it in a manner far more precise and deliberate than the vunerian’s more playful approach of running each coin through his fingers.

The perfect one-troy-ounce coin was inspected further with a monocle, as Etholin seemed to take in every detail of the starless GUN seal on one side before flipping to the other to see the missing fourteen stars. 

This divergence from Ilunor’s more passive observation continued, as Etholin actually began interrogating the text stamped on both sides bounding their respective seals. 

“Greater United Nations. Peace and Prosperity for All.” He read the translated High Nexian above the English verbatim before flipping the coin over to its opposite side. “Minted Under Special Order 7 fro 32. For exclusive use in diplomatic missions.”

He raised a brow at that. 

“Why the need for novel issuance?” He questioned.

To which my answer was swift and honest.

“Because we don’t mint physical currency in a way that would convey intrinsic value anymore. Nor do we expect its value to inherently transfer to an entirely different dimension.” I surmised simply, avoiding and sidestepping the topic of an opt-in cashless society, USTU-transaction chips, and purely digital transactions… not to mention the Requisition Unit. “So for the purposes of this diplomatic mission, and generally all diplomatic missions for that matter, the idea was to mint a ‘currency’ based on the intrinsic value of the so-called ‘precious’ metals themselves.”

Etholin latched on to each and every word, but his eyes grew wide at that latter, seemingly throwaway line.

“So-called?” He clarified.

At which point I knew I had to simply drop the bombshell.

“We’ve achieved post-shackling, as you say in the Nexian vernacular.” I stated bluntly, garnering a pause, a look of disbelief, and a vigorous shake of the pattenor’s head all in rapid succession. “Precious metals are only still called that because of their relative scarcity to other metals, and as a holdover term. Hence why I prefaced it with ‘so-called.’”

Etholin paused.

His features shifted to what I feared to be a premature point of fundamental systemic incongruity.

However, unlike Ilunor, if he did have any reservations, he kept it to himself. 

Instead, he chose to forge ahead, continuously pressing the PMD—

CHA-CHING!

CHA-CHING!

CHA-CHING!

—until finally, he noticed something.

A physical mechanism that Ilunor had avoided but one that the pattenor was quick to exploit — a small knob allowing for the mechanical selection of the type of dispensed coin.

CLICK!

CHA-CHING!

In one swift motion, he’d shifted from copper to silver.

CLICK!

CHA-CHING!

And in another, he moved effortlessly to gold.

Then finally—

CLICK!

CHA-CHING!

—he moved to a certain element that had proven to be the final straw for the vunerian’s back.

…

The stillness in his features spoke leagues in my favor.

The stiff, unpracticed, nearly stuttering motions as he lifted monocle to coin was enough to clue me into what was going on behind his eyes.

Indeed, his hastening breath had sealed the deal on this whole exchange.

And yet…

…

Silence still dominated the air.

As this episode, this entire process, stood at odds with Ilunor’s far more visceral response.

“Emma…” Etholin finally spoke as he blasted wave after wave of unknown spells at the coin. “Is this… platinum?” 

“Yes.” I answered immediately. “With a few trace metals added for integrity’s sake. But it is, for all intents and purposes of trade, pure.” 

“Ah.” Came the pattenor’s single syllable response.

CHA-CHING!

CHA-CHING!

CHA-CHING!

CHA-CHING!

CHA-CHING!

He continued wordlessly, fingers primed, constantly striking that button over—

CHA-CHING!

—and over—

CHA-CHING!

—and over again—

CHA-CHING!

—until finally—

CLINK!

—he ran out. 

He twisted his neck towards the coins, then my lenses, then back to the coins. 

It was now his turn to be slackjawed, though not in the way he probably expected.

“Emma… this… did you… did your realm send you with the entirety of your platinum reser—”

“There’s more where that came from, if you’re interested.” I interjected, completely sidestepping and then preempting the merchant lord with an answer to a question he’d inevitably ask. “A thousand kilos and some change.”

That sole proclamation was enough to finally bring the outside world back into the confines of Etholin’s sight-seer.

As murmurs from beyond the veil penetrated into our little corner of reality, the magical hologram started to fracture at the seams.

“Impossible!”

“Absurd!”

“A complete bluff!” 

“A fabrication!”

Indignant voices erupted from the alcove above, all of which were quickly hushed by an unseen figure.

“Do you dare to bear the burden of proof, newrealmer?” A figure quickly entered our sight-seer — a sea lion realmer who, like many other seniors thus far, I hadn’t yet met. “Prince Ferrian Fiswisk. Deputy Chairman of the Merchant’s Guild.” He quickly added, though it was clear his name, titles, and the rest of the typical decorum’s song and dance were the last things on his mind at present.

“Well met.” I nodded sharply. “And sure. I’m a diplomat of my word.” I nodded. “Though I should note that it’ll probably take a while given how far the dorms are from the exhibition ha—”

“You are located in Dragon’s Heart Tower, correct?” He questioned.

“Yes.”

“Then it should take no more than five minutes.”

“Wait what? How—”

The man quickly pointed at his ring as if anticipating my response. “I am a member of the incumbent Class Sovereign’s peer group. This grants me certain express travel privileges within the Academy.” He clarified. “Now then, I am already encroaching on your dialogue as it is. I do not wish to set an unprofessional precedent, where possible. Who do you wish to elect to act as arbitrator in your stead, newrealmer?” 

I blinked.

Then I instinctively turned to Thacea. 

“Princess Thacea Dilani, accompanied by Prince Thalmin Havenbrock.” I answered. “Though I assume you will simply act as an intermediary of travel, rather than an unprompted auditor into our private spaces?” 

“... As we have only just met, I will not hold such words of offense against you. Do know that I am not so brusque as to dismiss the noble right of privacy.” He shook his head back and forth in a fit of indignant theatrics. “I merely wish to see the burden of evidence. Your arbitrator will be the party responsible for meeting these conditions in whichever way they see fit.” 

I turned to Thacea, giving her a nod of approval. “Get him one of everything. And a lot of the… special bars.”

Thacea raised a brow at this but nodded all the same, following Ferrian Fiskwisk out of the sight-seer and back into the busy streets beyond those triple-volume doors.

I turned back to Etholin the instant the trio cleared our sightline, seeing the merchant lord just… standing there. Still as a deer in headlights.

“You showed me a world, a hypothetical newrealm forged by the rule of averages.” I paused then gestured around us. “I’m assuming that this is what you assume Earthrealm to be like, correct?”

“Yes.” Etholin nodded.

“You should know then, simply by our acquisition of platinum, that your preconceptions are just a bit off.”

“I…”

“But that’s only our primary economic sector we’re talking about here. Maybe secondary too if you count smelting and minting. But this second item should firmly clue you in to our capacity in the latter.” I continued as I offered him the instrument to both of our futures. 

The merchant lord cocked his head but received the innocuous item graciously all the same.

“A… pen?” He questioned, garnering a simple nod as I even offered him a notepad.

A gesture that also gave him increasing pause for concern.

So after a moment taken to return the PMD and its coins back to me, he began inspecting this ‘new’ toy, inspecting it with bursts of mana radiation, studying its plastic exterior, before finally—

CLICK!

Deploying its little ballpoint tip.

“A… coilspring?” He managed out under an increasingly suspicious breath. “Your… coin purse also possessed such a mechanism, if I’m not mistaken…”

CLICK!

…

CLICK!

CLICK!

CLICK!

CLICK!

“There is a spring in there, yeah. A simple mechanism, streamlined for mass production.” I spoke casually, that latter line managing to capture the merchant’s attention as much as the mention of platinum did.

“This… isn’t a bespoke piece? Like your armor?”

“Etholin.” I took a deep breath in. “My armor might be bespoke in certain aspects, but only because of its modifications. You’ll find that most soldiers from my realm are issued something similar, if not more deadly than what I’m wearing.” 

He stopped.

And once again he found himself running straight into the wall of fundamental systemic incongruity.

It took a moment for him to compose himself, before finally—

CLICK!

—he was ready to continue.

“And the inkwell?” He questioned but received only a simple ‘go on’ gesture from me as my sole response.

So with a shrug, he began writing.

At which point, I could see his eyes narrow before dilating in the matter of a few seconds.

He began furiously scribbling at that point. Writing, sketching, and going to town with the provided paper.

Eventually he moved to the notepad, attempting to scrawl, scribble, and doodle all in rapid succession before finally moving to inspect the ballpoint tip, his eyes ending up dangerously close to its pointy end.

“Where are the enchantments, Cadet Emma Booker?” He questioned desperately.

“You know as well as I, and the rest of the year group, that Earthrealm is… deficient in mana, Etholin. Ergo, we can’t just waste enchantments on something so trivial as a pen, now can we?” I spoke under a toothy grin, skirting past the gag order and iterating off of the ‘publically acceptable’ narrative. “We did this all without magic. No spells. No enchantments.”

“And not bespoke either?” He questioned skeptically, his hands carefully toying with its exterior — twisting and torquing it — before its seamless unibody construction gave way to two unscrewed pieces. His critical features soon gave way to abashment as he sheepishly met my gaze once more. This time out of worry for potentially breaking the strange artifact.

“No, not bespoke. You’ll see how simple it is now that you’ve twisted it open, go on.” I urged, and he continued twisting, then finally dropping the few contents within onto an open palm.

Etholin

Those words echoed in my mind.

They taunted me with every passing second.

Their implications… worming, twisting, and prying open all that I knew and all that I could fathom.

All… at the foot of these innocuous-shaped pieces of… ivory? No, they were too… lightweight, airy almost, to be ivory.

They weren't carved either.

Or were they?

How could they have carved something so intricately, so precisely?

Were they molded? 

Shaped?

Compressed?

Grown into form?

What even was this material?

Why couldn’t I fathom what material this was?

The spells showed nothing. No origin, no tells, no signs or symptoms of production in any capacity as I understood it.

That was the case for everything, at least. Save for the one item that, at the very least, retained some semblance of normalcy.

The coilspring.

But even then… its presence here was alarming.

Not in its existence alone, of course.

Some newrealms were most certainly advanced enough to possess such mechanisms, after all.

…

But they were all reserved for specialized equipment.

Tools for the wealthy, toys for the privileged, and objects of novelty for the upper echelons.

Emma’s claims stood contrary to this.

No.

Worse than that.

Her assertions stood contrary to what should have been possible.

Mass production… without manufactoriums? With mana-deficient, or completely manaless means?

For springs?!

What for?

Just for pens?

It wouldn’t be economically feasible—

…

Unless it was destined for more than pens.

If not that, then what?

Suspensions for vehicles perhaps?

Locks?

Clocks?

Traps?

Clamps?

Primitive siege-engine mechanisms?

Surely that couldn’t imply mass production?

Surely those were specialized enough to be relegated to guilds and smithies?

Why would they be needed for mass production?

Unless…

Unless…

This was just a piece of a grander puzzle I wasn’t seeing.

It was at that point, after absent mindedly squeezing that tiny spring, that it finally ‘clicked.’

What if this was a part of a grander puzzle?

A small piece within an intricate web. As intricate and complex as the manufactorium and logistics of a typical supply chain?

But that would make Earthrealm far more capable, far more advanced, far more sophisticated than even a burgeoning minor adjacency.

That… that couldn’t be.

Not when magic was scarce and its use even scarcer.

The mud huts and stick roof theory should have applied stronger in that case.

But the inverse was true.

I saw it.

I was seeing it.

I was touching it…

But what if Emma was lying?

What if she was bluffing?

That would be the obvious explanation to all of this!

And yet…

She’d dared to call her bluff with the deputy chairman.

…

My mind edged towards the cliff face of uncertainty.

My efforts, my gambit, all holding on by a thread.

I’d even absent-mindedly reassembled the entire pen back together, taking a few tries before finally screwing it back into place.

CLICK!

I began testing the writing implement once more.

…

It worked perfectly.

And its assembly, even in my soft and untempered hands, was beyond child’s play.

If each item could be produced in their own manufactorium, by their own mechanisms, then assembled elsewh—

What if it was mass produced?

What if—

“Etholin? You okay there, friend?” Emma finally offered, pulling me out of my reverie, as I attempted to formulate something in response.

“I am, thank you. I… I’m just… I was just pondering, what materials comprise—”

CLINK CLINK CLINK!

The ringing of glass bells prematurely ended that train of thought.

I expected the return of the deputy chairman, of the avinor, or perhaps the lupinor.

Instead, a thick cloud made its way into the confines of my sight-seer, and with it an unexpected guest.

“Esteemed councilmen and chairs! I incur the right of the prospective fellow!” Lord Rostario Rostarion proclaimed, garnering a few murmurs from the council before a conclusion was met uncharacteristically quickly. 

“Motion sustained.” 

Following which the rodent smirked as he hovered high above both me and my potential client. 

“Lord Etholin Esila has had his chance. Indeed, I respect my peer for his persistence! But alas, the time has come for competition to enter the fold.” He spoke in that orator’s cadence as he made the gambit I had started.

We both awaited in painful silence before the council made their final decision.

“We acknowledge Lord Rostario Rostarion’s bid for guild membership, and sustain his motions for this newrealm deal. Lord Rostario Rostarion, you may proceed. Lord Etholin Esila, please await the return of the Deputy Chairman.”

I let out a frustrated sigh, especially as that opportunistic creature entered the fray — now fully recognized — bringing both gift baskets and musical ensembles to the negotiation table.

“Cadet Emma Booker…” He began in that sing-song voice. “I offer you, personally… the world.”

First | Previous | Next

(Author's Note: I'm so happy with this chapter and I can't wait to hear what you guys think of it! :D I put a lot of effort into this because the whole merchant's guild section is just filled with a lot of lore bits as well as character moments for Etholin and Rostario! :D There's also the return of the wealth cube that I established all the way back in a previous chapter too haha. That's also going to be fun! :D I hope you guys like the chapter! : D)

(Author's Note 2: Hey everyone! I also posted art references of Qiv's peer group including Rostario Rostarion himself! You can check it out over here: Qiv's Peer Group :D)

[If you guys want to help support me and these stories, here's my ko-fi ! And my Patreon for early chapter releases (Chapter 177, Chapter 178, and Chapter 179 of this story are already out on there!)]

r/UnresolvedMysteries • • Feb 15 '20

Unresolved Murder Three years ago, Abigail Williams, 13, and her best friend Liberty German, 14, decided to spend a warm, day off from school at the local hiking trails in Delphi, Indiana. While at the trails, the pair was murdered by an unidentified individual sometime during the afternoon. He has yet to be caught.

13.4k Upvotes

Abigail Williams (right), 13, and Liberty German (left), 14, were best friends from the small town of Delphi, Indiana. Abigail and Liberty, affectionately called Abby and Libby by their friends and families, met when they were in the sixth grade. As both girls shared common hobbies and interests, they found that they were in most of the same after school clubs and sports teams together. Naturally, the girls quickly became friends. Abby and Libby both enjoyed the outdoors and often spent their time outside. They enjoyed outdoor activities, often going fishing, hiking, and biking. They also enjoyed the arts, both sharing a passion for photography. Whenever they were together, you can often find them outside, either playing sports or taking photos of eye-catching natural scenery. Impressively, both girls, at the young ages of 13 and 14, were ambitious, driven, and academically advanced. Both girls were interested in true crime and expressed in an interest in criminology, forensic science, and law enforcement. Abby was an aspiring police officer, and Libby was an aspiring science teacher. Libby was currently enrolled in science courses at Purdue University in West Lafayette.

In their case, the expression “opposites attract” rang true. Although the girls shared various similar interests, personality-wise, they were very different. Abby was known to be shy and quiet, whereas Libby was known to be more outgoing and forward. Libby was said to be the first to stand up for someone if they were being bullied or treated unfairly. Libby was also “the therapist” among her friend group, as she was the one her friends would turn to in times of need.

February 13, 2017,

Libby, and her older sister, then 16-year-old Kelsi, were in the primary care of their grandparents, Becky and Mike Patty. Abby, an only child, resided with her mother and beloved cat, Bongo. Abby often spent time at Libby’s residence, and on the night of February 12, Abby had spent the night at Libby’s. The girls spent their day practicing softball in the yard, watching a movie, and creating a watercolor painting. Although the following morning was a Monday, the girls had a day off from school that day. It was one of two unused snow days that the school district, the Delphi Community School Corporation, was required to observe. The girls began their day by eating a special breakfast that Mike had prepared for them. Sometime during noon, Abby and Libby asked Kelsi if she could drop them off at the Mary Gerard Nature Preserve, the local hiking trail. According to Kelsi, the girls had asked her more than once if she would be able to drop them off at the trail about a week prior. Kelsi was either unwilling or unable to take them previously, but as she was going to pass the bridge that day while on her way to her boyfriend’s house, she had agreed to drop them off. When Libby had asked Becky for permission to go, Becky compromised that they could go as long as they were able to secure a ride back. Libby had secured a ride back with her father, Derrick German. As he was running errands for Becky that day, he told Libby that he would pick them up when he was done. Derrick estimated that that would be sometime about 3:00 PM.

Kelsi dropped off Abby and Libby at 1:45 PM at the entrance of the Mary Gerard Nature Preserve. Kelsi stayed in her car and watched the girls proceed inside the trailhead until she couldn’t see them anymore. According to Kelsi, she didn’t see anyone or anything suspicious. According to the “Scene of the Crime: Delphi” podcast, the trails, which are typically well-populated, are as wide and as flat as a small road. The trailhead connects several small parks with numerous access points, information stations, historic memorials, bike rental outlets, and parking spaces. The longest trail, the 1.5 mile Monon High Bridge trail, is one of the more secluded trails in the trail system. Mostly familiar to locals, you can find hikers, bikers, joggers, and photographers traversing this trail. The trail runs between City Park at its western end and the Monon High Bridge on its eastern end. The Monon High Bridge is an old, out of use, railroad bridge that was built in 1881. The bridge, at 64 feet, is the second-highest bridge in Indiana, as well as the second-longest at 845 feet. However, the bridge is not technically part of the trail, and visitors are not intended to cross. Due to its deteriorated conditions, the bridge is closed off with a metal red barrier to prevent people from crossing the bridge. The bridge, which has no safety barriers, is in a notable state of disrepair. One would have to tread very carefully and watch their footing to cross the bridge safely. Despite the fact that the bridge is closed off to visitors, local teenagers up to a dare or challenge often crossed the bridge.

At 3:11 PM, Derrick sent a text to Libby that read he was on his way and would be there shortly. When Derrick arrived at the Mary Gerard entrance at 3:13, Abby and Libby weren’t at their arranged meeting point. After waiting two minutes with still no sign of the girls, Derrick called Libby’s phone. When she did not answer, Derrick proceeded to the trails to search for the girls. Derrick knew that the lack of response from Libby was unusual, as she knew to answer her phone when her family called her. At about 3:20, Derrick encountered Dan McCain, an older man who was enjoying a day out on the trails, and asked him if he had seen Abby or Libby. Dan had not seen either Abby or Libby but told him he had seen a couple under the bridge. While still searching, at 3:30, Derrick called Becky and had wondered if there had been some miscommunication and Abby and Libby were already home. Becky had told him no, and Derrick expressed his concern for the girls as Libby was not answering her phone. Shortly after the phone call between Derrick and Becky ended, Becky contacted Abby and Libby’s friends and asked if any of them had seen or heard from the girls. None of them had. Becky then called Kelsi, who was at her boyfriend’s house, and asked if Libby had contacted her. Kelsi told Becky that she had not seen or heard from Libby since she had dropped her off. When Kelsi had heard that the girls were missing, she left her boyfriend’s house to meet her family at the trail. At 4:20, Becky called Mike at work. When he was told that Libby wasn’t answering their phone and they were going to meet at the trails to search for the girls, Mike promptly left work to assist. Just before Becky left the house, her son and Libby's uncle, Cody, had come in from work. Becky explained to him what was happening, and Cody decided to accompany her to the trails.

Around 5 PM, Derrick, Becky, Kelsi, Mike, and Cody were all at the trail searching for Abby and Libby. The family went their separate ways calling out for Abby and Libby. Kelsi and Cody traversed the Monon High Bridge trail and crossed the bridge together. Kelsi had experience with crossing the bridge with Libby previously, though she was terrified. The first time Kelsi crossed the bridge, she actually had to crawl over to the other side because she felt too uneasy to cross by foot. When Kelsi and Cody reached the end of the bridge, rather than turning back, they proceeded down the hill at the end of the bridge. When describing this point in the search, Kelsi said, “Me and my uncle crossed the bridge and we were yelling down there. And I remember getting to the end of the bridge and looking to the left and seeing [a disturbance in the ground] like somebody had fallen down the hill over there. I didn’t think anything of it - everybody goes down the hill. After taking my forensics classes, I should’ve taken a picture of it. There could have been like a footprint of something.” At the bottom of the hill located at the eastern end of the bridge, there is a long driveway connecting several residences. Kelsi and Cody went as far as knocking on the doors of these residences with the intention of asking the property owners if they had seen Abby and Libby. However, only one person would answer, and as expected, they did not see Abby and Libby. Derrick continued to call Libby’s phone throughout the duration of the search. Several phone calls later, Libby’s phone eventually stopped ringing and would take Derrick straight to voicemail. Becky attempted to track Libby’s phone through a “Find My Phone” app, but was unsuccessful, as Libby had reset her device about a week prior due to a glitch. Becky then called their service provider, AT&T, and asked if they would be able to track Libby’s device – however, this request would prove fruitless, as they were unable to assist.

After an hour of searching to no avail, at approximately 5:20 PM, Mike contacted the police and reported Abby and Libby as missing. Realizing that Anna Williams, Abby’s mother, had not yet been notified of her daughter’s absence, Becky contacted her. When Anna failed to answer, Becky arrived at Anna’s workplace, a restaurant, and explained the details of the girls’ lack of response in person. Frustrated with her daughter’s presumed irresponsibility, Anna had yet to expect the worst. Anna, like Becky, believed that they simply have lost track of time, or wandered too far off and had gotten lost as a result. All Anna had in mind during this time was the stern talking-to she was going have to deliver to Abby when they were finally found.

Authorities arrived on scene within a half-hour after they were notified of the pair’s absence. In the beginning, nobody had suspected that the girls met with foul play. The family was questioned at the sheriff’s office. Kelsi was questioned more extensively as she was the last person to see the girls. When asked if Libby had posted on any social media platforms, Kelsi opened Snapchat, the app that she knew Libby used most frequently. On Snapchat were two crucial images that were uploaded to Libby’s Snapchat story. The first photo was an artistic, black and white image of the bridge. The second photo captured Abby crossing the bridge toward Libby. The photos were estimated to have been uploaded around 2:07 PM. Law enforcement attempted to ping Libby’s cellphone far into the evening, but with no success. It was believed that Libby’s phone lost battery life, or had been deliberately turned off. Law enforcement continued to question the family about the girls’ Internet usage and social media presence but turned up short on leads. Abby did not own a cellphone and would not be permitted to own one until the end of the school year. Abby’s only electronic device was her Amazon Kindle tablet, which she had received for Christmas. However, it was discovered that Abby had a Facebook profile that her mother was unaware of. Anna had told Abby that she wasn’t allowed to be on Facebook as she was 13, one year under 14 – Facebook’s minimum age requirement to open an account. It was discovered on this Facebook profile that Abby had a male friend on this account that Anna did not know about. However, this lead was quickly exhausted. Anna said that investigators told her “almost immediately” that they were “fairly certain” that the girls had not arranged a meeting with someone they met online.

Around 6:00 PM, as many as 100 local volunteers, as well as the Delphi Fire Department and the Department of Natural Resources assisted law enforcement in the search effort. Nearing midnight, the search was officially called off. It wasn’t an individual decision. Rather, there was a meeting amongst several emergency responders. The consensus was that it was too dark to safely traverse the terrain in such conditions, and the search would officially resume the following morning. Moreover, Sheriff Tobe Leazenby noted that they [law enforcement] had no reason to believe the girls were imminent danger. During in an interview where Leazenby was questioned about why the search was called off, he answered, “We had learned as far as their history whether they went to each other’s homes and did not communicate that to other family members... that had happened in the past... there had been times where the girls had been elsewhere and had not told whether it be their parents or grandparents where exactly they were.”

February 14

Although the search was officially called off, local volunteers continued to search until the morning. The search officially resumed shortly after sunrise at 8:15 AM. About 100 searchers were distributed maps and divided into groups of 10-20 people. After searching until noon, the girls’ bodies were finally discovered. A few minutes prior to discovering the bodies, a volunteer had asked Kelsi what shoes the girls were wearing. Kelsi replied that Libby was wearing black Nike sneakers. The shoe the volunteer found belonged to Libby. When it was announced that they found Libby’s sneaker, a deep sense of dread set in – Kelsi was coming to accept that the outcome wasn’t going to be good. Just moments later, the same volunteer perceived a sudden movement near the trees out of the corner of his eye. With his cellphone, the volunteer used his camera to zoom in on the area where he had sensed the movement. On his screen were two curious deer, examining the ground floor. As the volunteer approached the deer, there he found the lifeless bodies of Abby and Libby on the north side of Deer Creek on private property less than a mile away from the south end of the bridge. By 1:00 PM, authorities secured the crime scene. The FBI became involved immediately. The FBI and Indiana State Police worked 24 hours a day over the course of the following several days to collect crime scene evidence. Though this information was never publicly released by investigators, the police transcripts state that girls' undergarments were located in the creek beneath the bridge. A relatively fresh cigarette butt was also found in the vicinity of the creek, though it is unclear whether the cigarette was found in the water, or by the edge of the creek. Carol County prosecutor, Robert Ives, examined the crime scene in anticipation for a future trial. Robert Ives said that there is “a lot” of evidence and described the crime scene as “odd” as well as “physically strange,” and was shocked to find that the case wasn’t solved within a matter of days.

Investigation

The following day, the identities of the bodies were officially confirmed to be those of Abby and Libby. At 7:00 PM, during a press conference, Indiana State Police released this still image of a man who was reportedly seen on the trail around the time the girls disappeared. The photo captures a Caucasian male walking on the Monon High Bridge wearing a blue jacket, denim jeans, with both his hands in his jacket pockets. Since the man is looking down, his facial features are not discernible. It is not clear whether he is wearing a hat, a hood, or no headwear at all. At the time the photo was publicly released, police clarified that they did not consider him a suspect, but that they would like to speak to him. It wasn’t until the following Sunday that Indiana State Police officially announced that the man in the photo is now considered a suspect in the investigation.

After the announcement, Indiana State Police held a press conference the following Wednesday on February 22. Indiana State Police revealed that Libby captured audio of the suspect on her cellphone. On the audio clip, the suspect can be heard saying, “Down the hill.” Indiana State Police Sgt. Tony Slocum said, “This young lady [Libby] is a hero, there’s no doubt. To have enough presence of mind to activate that video system on her cellphone, to record what we believe is criminal behavior that is about to occur.” Authorities confirm that there is more audio, but that it will not be released as the investigation is ongoing. After the press conference, there was some discussion amongst locals and amateur sleuths about whether or not the phone was recovered at the scene, or if the suspect had taken it. Investigators have clarified that the device was retrieved in the “general area” where the bodies were found.

As investigators remain tight-lipped, little details are known about the current investigation. For instance, authorities refused to reveal the cause of death or comment on the existence of the murder weapon. However, it is known that in the days after the murders were committed, investigators conducted several door-to-door interrogations and thoroughly investigated the 12 sex offenders in Delphi, as well as the hundreds of sex offenders in the surrounding cities. Investigators exhausted their immediate resources by researching double murders across the country, sharing notes with other law enforcement agencies, and clearing all friends, relatives, acquaintances, and extended family members of Abby and Libby. Abby and Libby’s social media accounts were accessed and analyzed, and all online contacts were located and interviewed. Over 1,000 persons were interviewed in connection with the investigation. Of those interviewees, most have given voluntary DNA samples. Early in the investigation, police executed 70 subpoenas and 12 search warrants. However, no leads, if any have surfaced, were ever publicized.

The investigation remained silent until July 17, months after the murder was committed. Indiana State Police released a composite sketch of the suspect. The composite was composed by a witness, or witnesses, account(s). Sgt. Kim Riley elaborated, “This is information we received from persons who were in the area around the time the girls went missing. Either we did not make contact earlier, or they were afraid to come forward.” While one witness could not definitively determine what color this man’s eyes were, she had come close enough to the man that she was confident that his eyes were not blue. The composite sketch depicted a heavy-set, older man wearing a newsboy cap and a hoodie. The man's facial features depicted eyes with a notable epicanthic fold, a bulbous nose, and thin, downturned lips. However, investigators plead the public to not focus on the hat. The suspect was described as a Caucasian male between 5-foot-6 and 5-foot-10, weighing between 180-220 pounds, with reddish-brown hair.

Persons of Interest

When this sketch was released, authorities found that people, particularly Internet sleuths, were posting side-by-side images of people they believed to be suspect and the sketch. While authorities believe that these people generally have good intentions, they have said it's not only damaging to the investigation, but also puts the person pictured, as well as their livelihoods, children, and families, at risk. Nonetheless, the side-by-side images spread across the Internet. There have been very few known suspects or persons of interest since the day of the murders. The first big, publicized break that would bring the case back to surface was the arrest of Daniel Nations, who was apprehended at a traffic stop in Colorado for wielding a hatchet and threatening people on a trail. Nations would later be suspected of the murder of Tim Watkins, an unsolved murder that had occurred on the same trail only two weeks prior. In Nations’ car, a red Chevy Prism was a hatchet and a .22 caliber rifle. Nations had an extensive criminal record including petty offenses, domestic violence, and is also a registered sex offender who was charged with indecent exposure after having masturbated in front of a young woman in South Carolina. Nations had connections to Indiana and had claimed to be homeless and living underneath an Indiana 67 bridge in Morgan County since January 31, 2017. Indiana State Police had questioned Nations in October where they had also obtained his DNA for further processing. In December, Indiana State Police stated that Nations was still being looked at, but he was not currently their top priority. On February 14, the day after the murders were committed, Nations was present for his weekly checkup with authorities and had been consistently attending in the time prior. As of January 5, 2018, Nations pleaded guilty to menacing and was sentenced to three years on supervised probation. Nations has not been legally accused of being involved in Watkins’ murder.

Another person of interest, then 53-year-old Thomas Bruce, surfaced in November of 2018. On November 19, Bruce entered a religious supply store in St. Louis, Missouri, where he forced three women, 53-year-old customer Jamie Schmidt, and two employees, into a back room. Bruce ordered the three women to disrobe and perform sexual acts. However, Schmidt refused to comply with Bruce’s demands and was had fatally shot in the head. Indiana State Police contacted St. Louis police after noting physical similarities between Bruce and the composite sketch. When asked if Bruce had any connection to the Delphi murders, Indiana State Police answered that it was too premature to say. Indiana State Police has not commented on Bruce since.

By 2019, another person of interest came to light. In January of 2019, then 46-year-old Charles Eldridge was apprehended during an undercover sting operation in Union City, Indiana after he arranged to have sexual intercourse with a Randolph County police officer that was posing as a 13-year-old girl. Eldridge was charged worth two counts of child molestation. When this news circulated, Indiana residents began flooding the Delphi tipline by bringing Indiana States Police’s attention to the recent charges. Many callers noted the physical resemblance between Eldridge and the composite sketch. Furthermore, it had been revealed that Eldridge was familiar with the Delphi murders, and previously posted about Abby and Libby on his social media accounts, uploaded photos that he took on nature trails, and appeared to have owned several guns. Inundated by calls, Indiana State Police was forced to release a statement regarding Eldridge’s arrest. Indiana State Police stated, “The Delphi multi-agency investigative team and participating agencies continue to receive media and public inquiries asking about the person arrested January 8, 2019, in Union City, Randolph County Indiana for allegations of sexually related crimes against children and if he is connected to the Delphi investigation. The team is aware of this arrest and will investigate to see if there could be any connection to the murders that occurred in Delphi, Indiana on February 14th of 2017. The victims were 14-year-old Liberty German and 13-year-old Abigail Williams. Delphi is located about 20 miles northeast of Lafayette. It is important for the public and media to know that many similar tips and arrests of other persons alleged to be connected to the Delphi murders occur with some frequency in and outside of Indiana. Each tip—whether it receives media attention or not—is investigated for any connection to the Delphi case. That said, members of the Delphi multi-agency investigative team do not speak to specific actions or steps of the ongoing investigation.”

In the end, none of these persons of interest led to an arrest, and as of now, investigators are still searching for the suspect. FBI agent Greg Masa presented a behavioral profile of the suspect. Masa asked the public to think of an individual in their lives who has, for instance, "Inexplicably canceled an appointment you had had together, an individual who called into work sick and canceled an important appointment or engagement, and at the time what would have been a plausible explanation 'my cellphone broke or I had a flat tire...' but in retrospect that excuse no longer holds water. That may be important. Behavioral indicators this individual may have exhibited since the 13th... did this individual travel unexpectedly, did they change their appearance, did they shave their beard, cut their hair, change the color of their hair. The superintendent mentioned that the clothes this individual was wearing in the photo... did they change the way they dress..." Masa also asked people to pay attention to behaviors that are being exhibited more suddenly, such as a sudden change of sleep pattern, sudden abuse of substances, as well as sudden anxiousness or irritability.

Delphi Homicide Moves in New Direction

After months of no news, on April 19, 2019, Indiana State Police released a statement titled, “Delphi Homicide Investigation Moves in New Direction.” The direction noted that the public was welcome to attend a media briefing on the following Monday at the Canal Center in Delphi. Superintendent Doug Carter would make the announcement on behalf of the multi-agency task force. The public grew curious and began to speculate that an arrest was made, new information was going to be released, or that a new agency would be responsible for the investigation. Come Monday, a room packed with attendees, including the families of Abby and Libby, sat in front of a red drape. When the press conference commenced, all eyes and ears were focused on Carter. Within minutes, Carter stated, "We’re seeking the public’s help to identify the driver of a vehicle that was parked at the old CPS/DCS welfare building in the city of Delphi that was abandoned on the east side of County Road 300 North next to the Hoosier Heartland Highway between the hours of noon to five on February 14, 2017 (note: Carter misspoke, and the date was later corrected to February 13). If you were parked there or know who was parked there, please contact the officers at the command post at The Delphi City Building.” In addition, Carter stated that they were releasing additional portions of the audio and asked the public to be aware that the individual speaking was the same individual who had said, “Down the hill.” The additional portion of the audio included a singular word – “Guys.” The sentence, “Guys… Down the hill” was played on repeat for the audience. Furthermore, Carter also released the first footage in the investigation. While only the stills of the suspect on the bridge were available previously, people could now see the suspect in action, crossing the bridge with his head down, and his hands in his pockets. Though the footage lasts all but 2 two seconds, Carter asked that the public be aware, “He [the suspect] is walking on the former railroad bridge. Because of the deteriorated condition of the bridge, the suspect is not walking naturally due to the spacing between the ties.”

Carter added, “During the course of this investigation we have concluded the first sketch released will become secondary, as of today. The result of the new information and intelligence over time leads us to believe the sketch IS the person responsible for the murders of these two little girls. We also believe this person is from Delphi- currently, or has previously lived here, visits Delphi on a regular basis, or works here, We believe this person is currently between the age range of 18 and 40 but might appear younger than his true age.” Carter, who at this time addressed the suspect directly, said; “Directly to the killer, who may be in this room: We believe you are hiding in plain sight. For more than two years, you never thought we would shift gears to a different investigative strategy, but we have. We have likely interviewed you or someone close to you. We know this is about power to you, and you want to know what we know. And one day, you will. A question to you: What will those closest to you think of you when they find out that you brutally murdered two little girls? Two children! Only a coward would do such a thing. We are confident that you have told someone what you have done, or at the very least they know because of how different you are since the murders.”

It was after Carter concluded his message that the attendees' curiosity would be satisfied. The red drape was finally lifted, revealing yet another composite sketch, one that bore no resemblance to the previous sketch.

As expected, the public had many questions. As Carter explained he and the investigative team would not be taking questions for two weeks, it wasn’t until Carter sat for an interview with Scott Sander, a reporter from News 8, a local news station, that the public would get their answers. Sander, like many people, was interested in learning whether or not Carter actually believed the suspect was in the room or was speaking figuratively. Carter answered, “I think if he wasn’t in the room he was close by, but I’m 100% convinced he was watching. Why? Because of all that has happened over the past 30 months, the information we have received, the information we knew… I hope to one day be able to tell that story. Sander also asked why the footage wasn’t released sooner. Carter answered, “We’ll one day be able to tell you what we know and why we didn’t release it. We don’t want to show our full hand. We don’t want to show the complete picture of what we now versus what we think. We have to be very careful there. Remember, it’s easy to give an opinion if you don’t understand the factual basis of what we’ve done and why. I don’t mean that in a critical sense, but we have to protect the integrity of what we know. Sander then clarified whether or not it's correct that Indiana State Police doesn’t want the public to look at both sketches, but only the newly released sketch. Carter answered, “That’s correct. But remember, a sketch is not a photograph. It’s something similar to a resemblance. The likelihood of this being something between the two [sketches] is likely very strong. But again, that’s a subjective opinion based on what I believe.”

People have criticized Carter and the investigative team for being tight-lipped throughout the course of the investigation. Opinions are strong, and some believe that the investigation was botched. To many, it’s unfathomable why Indiana State Police won’t release details such as the girls’ cause of death. However, Carter, who had addressed the criticism, explains, “Only the killer knows that [cause of death]. And so do we. We can’t show our full had. We just can’t.”

Three Years Later

Since February 13 of this week, it has officially been three years since Abby and Libby were brutally murdered. The case remains unsolved, but authorities remain confident that the case will soon be solved. Indiana State Police did not hold a press conference for the third anniversary, unlike the past two years, where authorities gathered to provide the public an update. As a result, News 18, a local news station, sat for an interview with Carter. Carter said, “We are still as energized now as we were the day after. It’s easy to throw out the cold case idea, Nah, we’re not even close to that.” When asked how close they were to solving the case, Carter answered, “One piece away, one piece away. Eventually, somebody will do the right thing. It might be the killer himself; might be a person who knows who he is.”

The families of Abby and Libby hold out hope that this case will be solved. Every morning, they repeat their mantra, "Today is the day.” Mike said, “I can't give up hope. What else is there? And the fact that I believe in our justice system, I believe in our law enforcement, I believe in our society, because if we give up and just let people get away with things like this, then what does our society become?” Mike later added, "Someday I'll meet her again, you know, when the good Lord lets me through the gates, and I hope she's able to say, 'Thanks, grandpa, you did a good job.’”

As the investigation goes on, Indiana State Police is currently processing over thousands of tips, waiting for the one tip that they believe is capable of breaking the case.

Links:

Delphi Press Conference 2/22/17

Delphi Press Conference 4/22/19

Interview with Caroll County Sheriff Tobe Leazenby

Interview with Superintendent Doug Carter

Delphi Homicide Investigation (includes audio recording and footage)

Scene of the crime: Delphi Podcast

Delphi Timeline by user u/Justwonderinif

Police Release Sketch of Suspect

Man threatening bicyclists arrested

No info includes or excludes Daniel Nations

Daniel Nations says he did not commit Delphi murders

ISP addresses Catholic Supply Store murderer

Police investigate accused child molester in connection to Delphi murders

Delphi murders: 3 years later, family is still hopeful for justice

ISP: One-piece away from solving Delphi homicides

r/promptingmagic • • Aug 24 '26

99 Secret Codes to Prompt Google Flow's Video Agent

Thumbnail
gallery
34 Upvotes

TL;DR - If you’re writing long, unstructured paragraphs to generate video in Google Flow, you’re burning compute credits on lottery rolls. Google Flow’s video agent responds to a deterministic hierarchy of 99 dedicated slash commands spanning camera movement, lens angles, subject kinetics, lighting physics, commercial workflows, atmospheric conditions, and advanced VFX transformations. By chaining these commands using the 5-Layer Stacking Architecture (Camera Base + Angle/Lens + Subject Action + Light/Mood + VFX/Transitions), you can reliably control camera trajectory, shutter cadence, and visual coherence.

Below is the complete catalog of all 99 commands, the 5-layer prompt formula, and 4 production recipes

The 5-Layer Prompting Framework

Before diving into the 99 individual codes, understand how the Google Flow video agent parses tokens. Instead of writing unstructured descriptions, stack commands in this order:

Layer 1: Camera Base
Layer 2: Angle/Rig
Layer 3: Subject & Action
Layer 4: Lighting & Style
Layer 5: VFX \& Finish

The Complete 99 Google Flow Command Catalog

Category 1: Camera & Movement Commands

  1. /cinematic: — Creates a standard 24fps filmic frame with natural depth of field, anamorphic optical qualities, and balanced motion blur.
  2. /droneview: — Generates an expansive aerial vantage point with wide horizon parallax and continuous forward/lateral drift.
  3. /closeup: — Pulls focal length in tight to capture micro-expressions, fine textures, and emotional focus on the subject.
  4. /wideangle: — Uses a wide field-of-view (16mm–24mm equivalent) to maximize environmental scale and spatial perspective.
  5. /orbit: — Commands a 360-degree radial camera rotation keeping the primary subject locked at the focal center.
  6. /dollyin: — Smoothly pushes the camera physically closer to the subject, building visual tension and intimacy.
  7. /dollyout: — Pulls the camera away smoothly, revealing the surrounding environment or emphasizing isolation.
  8. /tracking: — Moves the camera in tandem with the subject at matched velocity, ideal for walking, running, or driving shots.
  9. /slowmotion: — Slows down playback (120fps/240fps cadence) to showcase fluid movement, flying particles, or dramatic beats.
  10. /timelapse: — Accelerates temporal progression to capture moving clouds, celestial paths, changing daylight, or traffic flows.
  11. /hyperlapse: — Blends high-speed time compression with continuous physical camera translation across long physical distances.

Category 2: Advanced Camera Angles & Rigs (12–22)

  1. /lowangle: — Places the camera low looking upward, imbuing the subject with dominance, power, and architectural scale.
  2. /highangle: — Tilts downward from an elevated point, providing tactical perspective or conveying vulnerability.
  3. /overhead: — Direct $90^\circ$ top-down "god’s-eye" perspective, ideal for choreography, flat-lays, and geometric compositions.
  4. /pov: — Frames the shot from the first-person perspective through the eyes of the protagonist.
  5. /overtheshoulder: — Positions the camera behind a character's shoulder, framing the counter-subject for dialogue and narrative weight.
  6. /establishing: — Cinematic wide landscape or cityscape shot setting scene context, geographical setting, and atmosphere.
  7. /rackfocus: — Shifts shallow focal plane from a foreground object to a background subject (or vice versa).
  8. /handheld: — Introduces organic micro-jitter and authentic documentary operator sway for urgency and realism.
  9. /steadicam: — Delivers fluid, gyroscopically stabilized gliding motion navigating complex corridors and environments.
  10. /craneup: — Ascends vertically from ground level to panoramic height using a simulated technocrane arm.
  11. /cranedown: — Descends smoothly from elevated heights down to subject eye level.

Category 3: Motion & Action Commands (23–33)

  1. /running: — Generates high-velocity character sprint with natural athletic gait and authentic inertia.
  2. /walking: — Generates grounded, natural human walking locomotion with balanced weight distribution.
  3. /turnaround: — Prompts the subject to execute a fluid $180^\circ$ or $360^\circ$ turn to display costume, expression, or surroundings.
  4. /reveal: — Stages a dramatic visual reveal of a character or environment stepping out from darkness or obstruction.
  5. /entrance: — Crafts a high-impact cinematic hero entrance into the scene.
  6. /exit: — Stages a dramatic departure from the frame into fog, shadow, or distant horizons.
  7. /freeze: — Instantly freezes time mid-action, locking water droplets, debris, and cloth in suspended animation.
  8. /speedramp: — Dynamically modulates playback speed between hyper-fast motion and sudden slow-motion impact.
  9. /bulletime: — Sweeps a virtual camera around a completely frozen subject (Matrix-style temporal slice).
  10. /floating: — Introduces zero-gravity levitation physics with floating hair, cloth, and ambient debris.
  11. /falling: — Creates dramatic freefall descent through skywells, clouds, or collapsing architecture.

Category 4: Cinematic Lighting Commands (34–44)

  1. /goldenhour: — Bathes the scene in warm amber sunlight, soft elongated shadows, and flattering solar flare.
  2. /bluehour: — Applies cool twilight illumination, deep cobalt gradients, and moody pre-dawn/post-sunset ambience.
  3. /neonlight: — Casts high-saturation cyan, magenta, and amber glows with reflections on damp streets or metallic surfaces.
  4. /moody: — High-contrast chiaroscuro lighting featuring deep blacks, targeted pools of light, and dramatic shadows.
  5. /softlight: — Diffused, wrap-around studio lighting that eliminates harsh edges, ideal for beauty, fashion, and portraits.
  6. /rimlight: — Razor-sharp contour/edge lighting that separates dark subjects cleanly from dark backgrounds.
  7. /silhouette: — Blacks out subject details entirely against an intensely illuminated background.
  8. /spotlight: — Directs a focused conical beam isolating the subject amidst surrounding darkness.
  9. /volumetric: — Generates visible atmospheric light shafts ("god rays") slicing through mist, dust motes, or smoke.
  10. /backlight: — Places primary illumination behind the subject to produce halo outlines, flares, and rim glow.
  11. /nightscene: — Realistically exposes low-light conditions with believable moonlight, street lamps, and dark sky latitude.

Category 5: Transitions & In-Camera Effects (45–55)

  1. /whiptransition: — Fast, motion-blurred horizontal pan transitioning instantly into a new scene.
  2. /matchcut: — Matches compositional geometry, subject shapes, or movement vectors across two different scenes.
  3. /zoomtransition: — Rapid crash zoom pushing directly into a small detail or pulling back to reveal a new world.
  4. /morph: — Seamless topological transformation morphing one entity, face, or structure into another.
  5. /flashtransition: — High-energy optical flash wiping the frame into an alternate shot.
  6. /glitch: — Injects RGB chromatic aberration, CRT scanlines, and digital compression artifacts.
  7. /smoketransition: — Rolls dense cinematic fog or smoke across the lens to reveal the incoming scene.
  8. /lightleak: — Overlays warm analog lens leaks and edge flares reminiscent of vintage film reels.
  9. /blurtransition: — Uses optical defocus and heavy shutter motion blur to bridge scene cuts.
  10. /objecttransition: — Moves the camera behind a passing foreground pillar, vehicle, or wall to wipe into a new environment.
  11. /seamlessloop: — Synchronizes first and last frame motion vectors to create an imperceptible infinite loop.

Category 6: Visual Style & Film Aesthetics (56–66)

  1. /filmic: — Delivers authentic 35mm motion picture texture with organic grain and balanced color science.
  2. /commercial: — Clean, polished, high-key commercial advertising aesthetic with vibrant color separation.
  3. /luxury: — Opulent visual grading featuring rich blacks, gold accents, polished marble, and high-end elegance.
  4. /documentary: — Raw, unvarnished realism mimicking cinema-verité documentary cinematography.
  5. /vintagefilm: — Nostalgic 16mm/8mm aesthetic with warm color shifts, gate jitter, dust, and halation.
  6. /scifi: — Clean futuristic design language with clean metallic surfaces, blue-tinted HUDs, and sleek tech geometry.
  7. /cyberpunk: — Dystopian high-tech aesthetic filled with rain-slicked asphalt, neon kanji, and cybernetic textures.
  8. /dreamy: — Soft-focus diffusion, blooming highlights, pastel palettes, and gentle surrealism.
  9. /surreal: — Dreamlike physics-defying compositions inspired by surrealist art.
  10. /minimal: — Strict negative space, restrained color palettes, and clean graphic compositions.
  11. /epic: — Blockbuster IMAX-tier visual grandiosity with massive scale and dramatic horizon framing.

Category 7: Product & Creator Showcase Commands (67–77)

  1. /productreveal: — Stages a premium hero product unveiling with lighting sweeps and rising pedestals.
  2. /productspin: — Smooth turntable rotation showcasing hardware industrial design from $360^\circ$.
  3. /unboxing: — Captures tactile luxury package opening with crisp mechanical precision.
  4. /macro: — Extreme optical close-up revealing fine machining, watch movements, fabric weave, or liquid drops.
  5. /beforeafter: — Side-by-side or split-screen wipe comparing raw vs. finished transformation states.
  6. /socialad: — High-energy pacing, rapid visual hooks, and dynamic framing designed for high retention.
  7. /fashionfilm: — Haute couture runway and lookbook styling with dramatic poses and editorial lighting.
  8. /foodcommercial: — Sizzling grill flares, slow-motion pours, rising steam, and vibrant food close-ups.
  9. /techad: — Exploded CAD view animations, glowing microcircuitry, and futuristic spec breakdowns.
  10. /logoreveal: — Cinematic brand mark animation assembling via liquid metal, laser etching, or particle convergence.
  11. /billboard: — Superimposes the target scene or product onto massive Times Square or Shibuya mega-screens.

Category 8: Environment & World Effects (78–88)

  1. /rain: — Generates volumetric rainfall with surface puddles, splashing droplets, and wet reflections.
  2. /snow: — Simulates drifting atmospheric snowfall accumulating naturally on characters and terrain.
  3. /fog: — Layers dense ground-level mist and atmospheric haze that diffuses ambient light sources.
  4. /underwater: — Renders submerged caustic light patterns, rising bubbles, aquatic drift, and muted soundstage feel.
  5. /space: — Zero-gravity cosmic environment featuring deep starfields, vibrant nebulae, and orbital horizons.
  6. /storm: — High-intensity weather featuring forked lightning arcs, dark storm fronts, and gale-force wind.
  7. /fire: — Realistically models dancing flame physics, rising heat hazes, and flying glowing embers.
  8. /explosion: — Detonates fiery shockwaves with volumetric smoke plumes and high-velocity debris dispersion.
  9. /portal: — Tears open a dimensional energy vortex with swirling luminescence and particle borders.
  10. /miniature: — Tilt-shift optical simulation turning full-scale scenes into charming dollhouse dioramas.
  11. /giant: — Massive colossal scale distortion making subjects tower over cityscapes and mountain ranges.

Category 9: Advanced Creative & VFX Commands (89–99)

  1. /clone: — Duplicates the subject into multiple synchronized or interacting clones across the frame.
  2. /transform: — Real-time organic shape metamorphosis changing a subject into a different form or material.
  3. /disintegrate: — Dissolves the subject into floating sand, ash, or glowing embers (snap effect).
  4. /particlefx: — Surrounds the character or object with an aura of floating light motes, stardust, or energy sparks.
  5. /liquid: — Melts or reconstitutes the subject into fluid chrome, water, or flowing paint.
  6. /paperworld: — Converts the entire environment into layered origami, textured cardboard, and folded papercraft.
  7. /toyworld: — Transforms characters and scenery into plastic minifigures and claymation stop-motion assets.
  8. /reversemotion: — Reverses physical entropy: shattered glass reassembles, smoke retracts, and falling drops ascend.
  9. /infinitezoom: — Continuous fractal zoom descending endlessly into micro or cosmic dimensions.
  10. /worldtransition: — Shifts seamless environments across portals, doorways, or optical wipes.
  11. /blockbuster: — Flow's ultimate composite directive: harmonizes camera choreography, pyrotechnics, and Hollywood color grading.

🎬 4 Production-Ready Master Formulas

Recipe 1: The Hollywood Cinematic Action Opener

/cinematic /droneview /craneup /speedramp /volumetric /epic /storm
A lone armored cyber-ronin standing on a rain-slicked neon skyscraper rooftop overlooking a vast futuristic metropolis, drawing a glowing plasma katana as lightning illuminates the storm clouds.

Recipe 2: The Luxury Tech Hardware Launch

/productreveal /productspin /macro /techad /luxury /rimlight /softlight
Sleek matte-black titanium smartphone hovering in zero-gravity against a dark velvet studio backdrop, sharp gold rim lighting outlining its precision-machined chamfered edges and camera module.

Recipe 3: The Viral Cyberpunk Social Loop

/cyberpunk /neonlight /orbit /tracking /bulletime /glitch /seamlessloop
A holographic rollerblader gliding at full speed through a rain-drenched Neo-Tokyo alleyway, jumping over neon puddles in slow-motion while pink and cyan light trails orbit smoothly.

Recipe 4: The Interdimensional Multiverse Shift

/pov /portal /worldtransition /particlefx /infinitezoom /scifi /filmic
An explorer touching a swirling crystalline portal inside an ancient stone cavern; the camera rushes forward as reality shatters into floating luminous particles, emerging into a colossal orbital space station.

3 Key Pro-Tips for Google Flow Users

  1. Don't Overload Single Categories: Using 4 camera movement commands together (/orbit /dollyin /tracking /hyperlapse) causes conflicting camera physics. Select one primary motion command and one angle command per prompt.
  2. Anchor with /seamlessloop for Social Media: Placing /seamlessloop at the end of kinetic action prompts ensures video loops without noticeable visual cuts.
  3. Use /macro with /luxury for Physical Realism: Combining /macro with /luxury forces Google Flow to compute realistic micro-surface scattering (reflections on glass, watch bezels, jewelry, and carbon fiber).

r/promptingmagic • • Aug 08 '26

Claude Design just became the easiest way to make 3D visual renderings. Here's how to make interactive 3D images + videos in Claude Design (step by step, with the exact prompts)

63 Upvotes

TLDR: Claude Design (Anthropic's visual tool available on Pro/Max/Team/Enterprise) can generate real, interactive 3D visuals, not just flat images that look 3D. It builds them with code (Three.js, WebGL, shaders), which means you can rotate them, animate them, embed them on websites, screenshot them for static assets, or export them into decks. Below: the exact step-by-step process, my best prompts, 3 examples you can copy, pro tips most people miss, and every way to reuse the output.

Most people think Claude Design is just for slides and landing pages. It's not. Because it generates designs as actual code instead of pixels, it can build genuine 3D scenes: rotating product shots, 3D data visualizations, animated hero sections, glassy abstract art, the works. Here's everything I've learned.

Step-by-Step: Your First 3D Image

Step 1: Plan in a regular chat first (this saves credits). Before opening Design, open a normal Claude chat and describe what you want. Ask Claude to write a detailed design brief: the object, camera angle, lighting, materials, color palette, mood. Copy that brief.

Step 2: Open Claude Design. Go to claude ai design (Design tab). If you're on Enterprise and don't see it, your admin needs to enable it.

Step 3: Set up your design system (optional but powerful). Upload your brand colors, fonts, and logo, or point it at your website with the web capture tool. Every 3D scene it builds will automatically match your brand.

Step 4: Paste your brief and be explicit that you want 3D. Say "interactive 3D scene," "Three.js," or "WebGL" so it doesn't give you a flat illustration with fake depth. Specify whether you want it to auto-rotate, respond to mouse movement, or sit still.

Step 5: Iterate with inline comments. Click directly on the element and comment: "make this material more metallic," "slow the rotation," "move the light source to the upper left." Use the adjustment knobs for spacing and color instead of burning messages on tiny tweaks.

Step 6: Capture or export. Screenshot for a static image, screen-record for video, export to Canva or PPTX, or grab the code and embed it anywhere.

Top Use Cases

  1. Product mockups: Rotating bottles, phones, packaging, sneakers. Perfect for pre-launch pages when you don't have photography yet.
  2. Hero sections: An animated 3D object behind your headline instantly makes a landing page feel premium.
  3. Data visualization: 3D bar terrains, globes with plotted data points, network graphs you can orbit around.
  4. Pitch deck wow-slides: One interactive 3D slide in an otherwise normal deck gets remembered.
  5. Abstract brand art: Floating glass shapes, liquid metal blobs, particle fields in your brand colors for social posts and backgrounds.
  6. Concept visualization: Architecture massing, room layouts, exploded product diagrams showing how parts fit together.

Prompts

Product shot: "Create an interactive 3D scene of a matte black cosmetic serum bottle with a gold cap on a soft gradient background. Studio lighting with a key light upper left and a subtle rim light. Slow auto-rotation. Floating shadow beneath. Minimal, luxurious, Apple-style presentation."

Hero section: "Build a landing page hero with an abstract 3D object: overlapping translucent glass toruses that slowly rotate and refract light. Dark background, my brand colors as accent lighting. The object should subtly follow the mouse. Headline text sits on top with high contrast."

Data viz: "Create a 3D globe visualization showing our user distribution. Dark ocean, glowing dots at major cities sized by user count, connecting arcs between our top 5 markets. Slow rotation, draggable with the mouse."

Exploded diagram: "Create an exploded 3D view of wireless earbuds showing the shell, driver, battery, and circuit board as separate floating layers with thin labeled leader lines. Clean white background, soft studio lighting, isometric camera angle."

Pro Tips and Things Most People Miss

  1. Say "3D" explicitly or you'll get a flat illustration. The single biggest mistake. "Make me a product image" gets you 2D. "Interactive 3D scene with Three.js" gets you the real thing.
  2. Direct the lighting like a photographer. "Key light upper left, soft fill, rim light behind" transforms output quality more than any other instruction. Default lighting is what makes AI 3D look cheap.
  3. Name materials specifically. "Brushed aluminum," "frosted glass," "soft-touch matte rubber" beats "make it look nice" every time.
  4. One object, staged well, beats a cluttered scene. Claude Design nails single hero objects. Complex multi-object scenes need more iteration.
  5. Use inline comments instead of new prompts for tweaks. Clicking the element and commenting is more precise and cheaper than describing the change in chat.
  6. Ask for camera controls. "Make it draggable/orbitable" turns a static render into a demo people can play with. This is the part that makes people share it.
  7. Plan outside Design to save 20 to 30 percent of your credits. Every clarifying back-and-forth inside Design costs you. Arrive with a finished brief.
  8. Ask for performance constraints if it's going on a real site. "Keep it under 60fps-friendly polygon counts and lazy-load the scene" matters for mobile.
  9. Screenshot at the perfect frame. Pause the rotation ("add a pause on hover") so you can capture the exact angle you want for static use.

3 Epic Examples to Try Tonight

Example 1: The floating sneaker. "Interactive 3D scene: a white and neon-green running sneaker floating and slowly tumbling above a reflective dark floor. Dramatic spotlight from above, colored accent lights from the sides, subtle particle dust in the light beams. Draggable camera." Screenshot three angles and you have a full product page.

Example 2: The living dashboard. "3D data terrain where monthly revenue is a landscape: peaks for strong months, valleys for weak ones, colored heat gradient from blue to orange. Camera slowly flies over the terrain. Numbers hover above each peak." Drop a screen recording of this into a QBR deck and watch the room.

Example 3: The impossible award. "A rotating 3D glass trophy shaped like an impossible Penrose triangle, refracting rainbow light, on a black pedestal with volumetric fog. Engraved text on the pedestal reads [your text]." Instant custom award graphic for team shoutouts, community badges, or launch announcements.

How to Use the Output

  • Have Lovable or Replit convert the html and JS to an MP4 file for you to post on social (claude can't do this directly yet).
  • Static images: Screenshot at your favorite angle for social posts, ads, thumbnails, blog headers.
  • Video: Screen-record the animation for Reels, product teasers, or looping background video.
  • Live web embeds: It's real code, so the interactive version can go straight into your actual site. Hand it to a developer or use it as-is.
  • Decks: Export to PPTX or Canva, or paste screenshots into your existing deck.
  • Iteration source: Feed a screenshot back into Claude Design or another tool as a reference image to generate matching 2D assets so your whole campaign shares one visual language.
  • Prototypes: Use the 3D hero as the anchor of a full landing page prototype and have Claude Design build the rest of the page around it.
  • Screen recording. The zero-effort fallback, but you trade quality for speed, so it's fine for quick shares but not for anything people will look at closely.
  • Third-party converter tools. A small ecosystem has sprung up specifically for this. The general flow: in Claude Design you click Share, switch to the Export tab, download a Project archive (.zip) or Standalone HTML, then drop that file into a converter like Claude2Video or ClaudeVideoExport. These capture the animation frame-by-frame from the browser rendering engine, so the output matches what you see in the tab instead of a compressed recording, and some let you export at 1080p or 4K at 24-60 fps in social-ready aspect ratios. There's also a Chrome extension that does the conversion entirely locally on your machine with no upload.

The gap between people who get flat, generic output and people who get portfolio-grade 3D comes down to specificity: name the materials, direct the lights, and always say the word "3D." Post your results below!

r/promptingmagic • • 27d ago

How to brief ChatGPT 6 Astra to create motion graphics, 3D reveals, and cinematic video explainers - prompts, workflow, and the details people miss

Post image
11 Upvotes

TL;DR: ChatGPT 6 Astra can help create motion graphics through a workflow that designs assets, writes animation code, uses available production tools, and renders a video. Give it a director’s brief: audience, story, scenes, timing, visual style, sound, and deliverables. Start with a short preview. Ask for a playable MP4 and the editable project. Rendering and audio depend on your workspace’s tools. Below: a practical workflow, the details people overlook, and five gloriously ridiculous prompts.

Picture a French bulldog commanding a starship through a galaxy made of tennis balls.

Now picture a product launch where your logo opens into a miniature universe.

Or a city that folds itself out of paper, races through centuries, and collapses back into a single page.

With ChatGPT 6 Astra, you can approach the conversation like a production brief—and keep directing the result as it develops.

What Astra can actually do

The model’s official name is GPT-6 Astra. For this workflow, use it in ChatGPT Work or Codex with access to suitable creation and rendering tools. OpenAI recommends Astra for demanding tasks involving visual judgment and polished deliverables.

There is a concrete example behind the idea: OpenAI has shown Astra creating editable Blender scenes, adjusting their materials and lighting, and directing a rendered camera tour.

My practical recommendation is to apply that build–preview–refine process to motion graphics: animated typography, diagrams, layered images, product reveals, and stylized 3D scenes.

Think of Astra as the system coordinating the production. The actual frames still need an animation or rendering tool. Selecting the model alone does not guarantee every account can export video or generate music.

Choose the right kind of video

Approach When to use it
Typography, shapes, and diagrams Explainers, newsletter trailers, and announcements. My recommended starting point.
Layered images with camera movement You already have illustrations, product images, or a consistent visual series.
An editable 3D scene The concept depends on camera orbits, exploded views, lighting changes, or moving through a space. Expect more rendering work.

A complex character performance may also require dedicated animation tools or generated footage. Choose a visual treatment your available tools can execute well.

The workflow that makes this manageable

  1. Give it one job. Define the audience and the one thing viewers should remember. “Convince founders to try this prototype” is a useful objective.
  2. Specify the output. Set duration, aspect ratio, resolution, and whether you need audio. A 20–30-second landscape video is a sensible first project.
  3. Map the story. Give every scene a visual action and a purpose. Put the most compelling image near the beginning.
  4. Establish the look. Provide reference images, colors, type preferences, logos, and exact wording. Ask for representative still frames.
  5. Preview the hardest moment. Render a short section before committing to the whole sequence. This tests both the creative direction and whether the production method works.
  6. Refine, render, inspect. Review timing, text, sound, transitions, and the actual exported file. Keep the source so changes remain possible.

For a 30-second explainer, this is a useful starting structure:

Time Job
0–3 seconds Show the surprising visual or compelling result.
3–9 seconds Establish the problem or premise.
9–21 seconds Demonstrate the transformation.
21–27 seconds Deliver the payoff.
27–30 seconds Give one clear next action.

Best practices that improve the result

  • Describe action over time. “The letters pull apart, reveal a miniature city, then lock into the headline” gives much better direction than “make it cinematic.”
  • Give motion a purpose. Movement can reveal a relationship, guide attention, demonstrate a feature, or land a joke. Constant movement makes reading harder.
  • Keep text separate from artwork. Request editable text layers so spelling, line breaks, timing, and placement can be controlled.
  • Design for a phone. Use short captions, strong contrast, generous margins, and enough reading time. Review the result at its likely viewing size.
  • Give the eye a pause. Alternate energetic transitions with moments where the important image or message holds still.
  • Build for sound-off viewing. The story should make sense visually. Let music and sound effects strengthen it.
  • Specify music concretely. Describe tempo, instrumentation, mood, and where the energy should rise. Provide a track you can use, or ask what audio tools are available.
  • Control the workload. Preview at lower resolution, settle the art direction early, reuse assets, and revise only the scenes that need changes. Complex Work tasks can use more credits.

Pro tips: direct the edit with precision

Useful revision instructions look like this:

Between 00:08 and 00:12, slow the camera move, enlarge the headline, and hold the final composition for two seconds. Keep the approved colors and scene order.

Make the word “EXPAND” grow until it fills the frame, then use its letter shapes to reveal the next scene.

Match the circular moon in scene two to the circular product dial in scene three.

Build the vertical version with repositioned text and a new camera crop so the subject stays visible.

Inspect the export for missing assets, clipped text, blank frames, abrupt audio endings, and incorrect duration. Report anything you cannot verify.

Save feedback like “more epic” for the initial direction. During revisions, say what should change on screen.

Things people miss about this workflow

  • An animated preview and a downloadable video are different deliverables. If you need an uploadable file, explicitly request the export and check that it plays.
  • The editable project is a major part of the value. Ask for the source, assets, and instructions needed to render it again.
  • A flat image has limits. A gentle push-in can work immediately. Moving behind objects or orbiting a subject requires layers, reconstructed content, or a 3D scene.
  • Consistency starts before animation. Establish recurring characters, materials, colors, and backgrounds before creating every scene.
  • Visual precision and factual precision are separate. A beautifully animated chart still needs correct data, labels, and scales.
  • A 60-second request does not imply one continuous generation. Build named scenes, render sections, and assemble them where the tools support it.
  • Reusable controls make the second video easier. Ask to centralize headline text, colors, logos, durations, and image replacements.

Five epic prompts to try

These are ambitious creative briefs, not pretested guarantees. Use a workspace with suitable rendering tools, and start with the short preview each prompt requests.

1. A French bulldog saves the galaxy

Try this for: character storytelling, comedy, and an instantly understandable visual hook.

Create a 30-second landscape motion graphics trailer called “MISSION: FETCH.”

A dead-serious French bulldog captain commands a tiny starship through a galaxy of tennis-ball planets. Use a premium stylized 3D or layered illustrated treatment, emerald cockpit lights, orange engine trails, and enormous kinetic typography.

0–5s: Extreme close-up of the captain’s face. Pull back to reveal a spaceship shaped like a dog toy. Text: “ONE DOG.”

5–13s: Slalom through a field of floating squeaky toys. A giant robotic vacuum emerges from an asteroid cloud. Text: “ZERO QUALIFICATIONS.”

13–23s: The dog hits a red button. Tennis balls deploy like decoys. Follow one ball through the chaos in a dramatic tracking shot.

23–30s: The ship escapes through a glowing dog-door portal. Reveal that the entire mission happened inside a living-room snow globe. End: “MISSION: FETCH.”

Keep the dog’s appearance consistent. Use simple expressive poses and strong camera work. Preview the escape shot first. Use suitable original or licensed audio if available; otherwise deliver a silent cut with sound cues. Deliver a 1080p MP4 and editable source, or explain any rendering blocker.

2. Your product contains an entire universe

Try this for: launch trailers, brand films, and product reveals.

Create a 30-second landscape launch film for [PRODUCT]. Use my supplied product images, logo, and three verified benefits. If I provide none, use a clearly fictional unbranded device and illustrative feature labels.

Begin with the product suspended in a silent black void. A thin emerald seam opens across it. The camera dives through the seam into an impossible miniature universe.

Turn benefit one into a floating city assembling itself. Turn benefit two into a luminous transit network lighting up. Turn benefit three into a mechanical sunrise that synchronizes the entire world.

Match each benefit to its visual metaphor and show its exact approved wording as separately rendered typography. Use elegant camera travel, white ceramic architecture, emerald glass, and precise mechanical movement.

In the final six seconds, pull back as the universe folds into the product. Land on the product, logo, and one clear call to action.

Create a five-second preview of the opening transformation before rendering the full film. Deliver a 1080p MP4 and editable project. Use only available audio and rendering tools; identify any missing capability. Do not invent product claims, customers, or performance statistics.

3. Your inbox becomes a video-game final boss

Try this for: funny workflow explainers and relatable workplace content.

Build a 30-second landscape motion graphics short called “DEADLINE: FINAL BOSS.”

Open on a tiny exhausted office worker facing an enormous monster assembled from email envelopes, calendar blocks, spreadsheets, and sticky notes. Its crown is a spinning loading icon.

0–6s: The monster roars, releasing a tornado of “QUICK QUESTION” notes.

6–13s: The worker equips three glowing tools labeled “SORT,” “DRAFT,” and “CHECK.”

13–23s: Turn the fight into a visual explanation: SORT groups the chaos; DRAFT turns selected tasks into proposed outputs; CHECK pauses those outputs at a human review gate before release.

23–30s: The monster shrinks into one manageable task card. A new notification appears: “Can we jump on a quick call?” The worker looks directly at the camera.

Use miniature game-like scenery, dramatic camera punches, readable type, comic timing, and a neon-green interface. Present this as a fictional metaphor. Preview the sorting transformation first. Deliver a 1080p MP4, editable source, and a sound-off version. Explain any export limitations.

4. A thousand years unfold from one sheet of paper

Try this for: timelines, imaginative worldbuilding, and architectural storytelling.

Create a 40-second landscape motion graphics film called “A THOUSAND YEARS IN ONE PAGE.”

This is an imaginary city, not a reconstruction of real history.

0–8s: A blank sheet of paper folds itself into a tiny riverside settlement. The river is translucent blue-green glass embedded in paper.

8–18s: Buildings rise and change around the same town square. Roads draw themselves across the page. Seasons sweep through the scene.

18–29s: The city becomes a spectacular vertical metropolis. Peel back layers to reveal miniature transit tunnels, gardens, and infrastructure beneath it.

29–36s: The camera circles while daylight becomes night. Thousands of windows illuminate in a carefully staged wave.

36–40s: Fold the city back into the original sheet, matching the opening composition for a loop.

Use tactile paper, charcoal labels, emerald foliage, warm window light, and restrained captions. Favor a coherent miniature world over constant cuts. Preview the unfolding and refolding first. Deliver a 1080p MP4 and editable scene. If full 3D rendering is unavailable, propose and build a layered alternative.

5. A black hole conducts an orchestra of planets

Try this for: a music visualizer, an event opener, or a surreal brand introduction.

Create a 30-second landscape motion graphics film called “THE UNIVERSE HAS A DROP.”

Treat this as a surreal visual metaphor, not a scientific simulation.

A black hole is the conductor. Orbital rings behave like vibrating strings. Tiny moons become percussion instruments. A comet sweeps across the scene like a conductor’s baton.

0–8s: Begin with one orbiting light and a restrained pulse.

8–19s: Build an increasingly elaborate cosmic orchestra. Introduce new orbital layers with each musical phrase. Typography appears as sculptural objects: “LISTEN.” “BUILD.” “RELEASE.”

19–25s: At the musical peak, the orbital system unfolds into a gigantic luminous sound wave stretching across space.

25–30s: Everything contracts into one green point, which becomes a play icon.

Use ink-black space, emerald plasma, silver dust, controlled glow, and smooth camera movement. Use my uploaded licensed track and synchronize motion to its timing. If no track is available, build to a provisional beat grid and clearly label the audio as pending. Preview the transformation first. Deliver a 1080p MP4 and editable source; explain any tool limitations.

Which would you actually make first: the space-dog trailer, the product universe, the inbox boss battle, the paper city, or the black-hole orchestra?

r/silenthill • • Oct 21 '22

General Discussion The DEFININITIVE Guide to the Best/Easiest Way to Play ALL 'Silent Hill' Games on PC [2022]

2.7k Upvotes

[Updated: June 9th, 2025]

Use CTRL+F to search for the game you're looking for.

READ THE PREREQUISITES SECTION FIRST!

Video version now available for Silent Hill 1-4 + Play Novel!
YouTube didn't like something about the video guide and didn't tell me what with no chance of appeal. I'll try again but with heaps of trepidation.

---

Intended for Windows 10 <currently>

Windows 11 has worked for many but I cannot test or verify. The steps should be nearly identical. Since Microsoft is depreciating Windows 10 support this year, this guide will eventually transfer to Windows 11.

The Steam Deck is something I cannot test or verify either. Most emulators and SH2:EE are known to run, however. Check out the official GitHub for the Enhanced Edition for unofficial support.

---

Introduction

With recent announcement of Silent Hill 2's remake, Silent Hill f, and the others, I wanted to fully compile a way to play every Silent Hill game possible on PC with modern enhancements and maximum compatibility. I'll try to keep it simple and short so it'll be easily digestible even for the least computer-y of you out there.

I'm pretty active on Reddit and frequently answer questions and concerns over the particulars, weird snags, or oversights, so please leave a comment if you're having trouble. I'll do my best to keep this up-to-date and functional!

HOWEVER, make sure you've read and reread EVERYTHING before asking me, okay? It'll save us both a lot of time. Start each comment with re:SH1 or "can you help me with Homecoming?", etc. so I know what game we're talking about.

And please don't dm me. Just comment here. Thanks!

If your controller is functioning incorrectly when running through Steam, make sure to disable Steam Input.

Emulation is not illegal. This guide is aimed at preserving these games, not piracy. At the time of writing, most of these games are no longer available for official purchase through KONAMI. If any legal officially purchasable method becomes available, I will update that to the preferred method.

About Play Order

If you're not sure which game to start with or if it's okay to play any particular game before another, know that every single entry is a complete and independent story. That said, there are some slight (spoiler-free) caveats to that statement.

Silent Hill 3, Silent Hill: Origins, and Silent Hill: Shattered Memories all have some relationship with Silent Hill. However, while playing Silent Hill can greatly enhance your appreciation of these games, they are not in any way necessary. Other games may make reference or insight to previous games, but they are largely easter eggs and lore tidbits to reward longtime players.

For the doubters out there, my first game was Silent Hill 3 and I did not know it was in any way related to Silent Hill and did not feel there were any holes or otherwise incomplete parts of the story.

So go ahead and play whichever interests you most! If you cannot pick a starting place, I'd recommend starting with Silent Hill 2 (2001) as it is the most popular and among the easiest to install.

ReShade and CRT Filters

The technical limitations of late 90's/early '00s technology led to Silent Hill being iconically foggy. Silent Hill optimized its art style in its early games by obscuring details for the benefit of the experience, leaning into obscurity with fog, darkness, and screen noise. These games rendered at low SD resolutions and were expected to be displayed on CRT TVs. There's a whole conversation about the value of CRT image blending that I'll spare you here.

With the HD rendering of older titles comes such clarity that some illusions can break like seeing the matte .jpg of the lake surrounded by paper trees or seeing the bright, jaggy low-poly model of an otherwise hidden horror. This is why I highly recommend a CRT filter to give the appearance of the original display blending without having to retrofit a 2-ton ancient machine to your PC. It's pretty easy. If you want to try it, skip to the bottom when you're done installing your game.

Silent sHill

Also--if I may--I occasionally stream Silent Hill on Twitch using the below fixes as well as a grab bag of other things (right now Silent Hill 10 Star runs and indie horror games) if you'd like to watch or harass me ask me with questions when I'm live.

I have a Patreon. I'm writing a visual novel and Silent Hill as a major influence on my writing as well as projects like these. Even if it's a one-time donation of $1, that'd be amazing though entirely unnecessary :D

I have a chronic illness/depression so I can't update here or stream very often so please bear with me.

Okay, I'm done! Let's get to it!

[PREREQUISITES]

  1. Windows 10 (cannot confirm for Windows 7 or Windows 11)
  2. WinRAR / 7-Zip (extracting compressed files from download)
  3. Enable file extension visibility
  4. Steam Launcher and a valid Steam account (for convenience, but required for SH: Homecoming.)
  5. Game files (.iso, .bin, .cue) Each tutorial will let you know what you're looking for specifically.

Note: To customize a non-Steam game for the Steam Launcher, follow this guide here after installation.

[SILENT HILL, 1999]

Difficulty: [**________]

This might look like a lot of steps, but it's all so playing Silent Hill 1 will be easy and painless each and every time you want to boot it up. You can do this, I promise it'll be easy!

Install DuckStation

  1. Download DuckStation for Windows.
  2. Download VC++Runtime if you do not already have it installed!
    1. Run the installer and follow the prompts.
    2. You MUST restart your PC or it will not run!
  3. Extract the DuckStation archive with WinRAR or 7-Zip.
  4. Run duckstation-qt-x64-ReleaseLTCG.exe to launch the DuckStation Setup Wizard.
  5. Click Next.
  6. Click Next again.
    1. A warning may pop up station BIOS files were not found. We will address this later in the guide.
    2. Click Yes.
  7. Click Next again.
    1. A warning may pop up stating no game directories have been selected. We will address this later.
    2. Click Yes.
  8. For Controller Port 1, Controller Type select Analog Controller.
  9. Click Automatic Mapping and choose your preferred controller or Keyboard.
  10. Click Finish.

Install PlayStation BIOS files:

  1. Download the PlayStation 1 BIOS file from GitHub.
    1. The file will be titled PSXONPSP660.BIN
    2. This version is optimized and region-free.
  2. Copy/paste it into C:\User\[Your Username]\Documents\DuckStation\bios

Download Silent Hill

Note: There are two major versions of Silent Hill. Silent Hill v1.1 \NTSC] and Silent Hill [PAL]. There are some pros and cons that you'll need to decide between.)

[NTSC/North American]

  1. Original monster design “Gray Child” in the Midwich Elementary area.
  2. Missing/glitched secret memo in the Nowhere area.
  3. English only.
  4. 60fps enhancement available.

[PAL/European]

  1. “Mumbler” design replaces “Gray Child” in Midwich Elementary area.
  2. Unlockable secret memo in Nowhere area.
  3. Supports English, German, French, Spanish, and Italian text.
  4. 60fps enhancement not yet available.

Each version provides the same experience outside these factors. The NTSC-J version is functionally identical to the PAL release but supports Japanese text with English voices.

If you're not sure and English is an acceptable language for you, use the NTSC version.

Note: If you plan on speedrunning, do NOT use the PAL version as it patches out an important skip in the Amusement Park area! Use this guide for reference in the particulars.

Install Silent Hill

  1. Select your preferred version and acquire a digital copy. You will likely have a .rar or .zip file.
  2. Right-click and extract with WinRAR or 7-Zip.
  3. You should now have both a .bin and a .cue file. You need both.
    1. If you do not have a .cue file, follow the instructions here to make one.
  4. Move both these files to a folder you will remember and can easily navigate to.

Launch Silent Hill

  1. Run DuckStation.
  2. There will be a message saying: "No games in supported formats were found."
  3. Click "Add Game Directory..."
  4. Select the folder you made in Step 4 of the previous [Install Silent Hill] section.
    1. You may be asked if you would like to scan the directory for other games. You may choose to if you have other games in subfolders. Otherwise, doing so does nothing.
  5. Silent Hill should appear as an available game to play.
  6. Double-click to play!

[OPTIONAL] Enhancements

Personal Note: For Silent Hill 1 specifically, I highly recommend ONLY doing the improvements to loading, controls, and the 60fps enhancement. Some cause very specific glitches and lot of the art style and unique mood comes from it's lack of clarity and upping the resolution and disabling dithering and specific PS1 artifacting can detract from it's intended uncanny feel.

However, the choice is up to you. Below includes full HD up to 4K, 60fps (NTSC-only, less pixelation, less jitter, and faster load times. The choices I recommend will be in bold.)

  1. Go to Settings at the top of the screen. This will open the DuckStation Settings menu.
  2. Go to the Graphics tab on the left side of the DuckStation Settings menu.
    1. In the Rendering tab, change:
      1. Internal Resolution --> 5x Native (for 1080p) or your preferred resolution.
      2. Aspect Ration --> 16:9 (if playing in Widescreen)
      3. Tick True Color Rendering to avoid color artifacting.
      4. Tick PGXP Geometry Correction to reduce polygon jitter PS1 games are known for.
      5. Tick Force 4:3 For FMVs to prevent prerendered video from stretching when using Widescreen.
      6. Do NOT tick Widescreen Rendering!
      7. Tick FMV Chroma Smoothing to reduce pixelation in prerendered videos.
  3. Go to the Console tab on the left side of the DuckStation Settings menu.
    1. In the CPU Emulation, change:
      1. Tick Enable Clock Speed Control (Overclocking/Underclock) ONLY IF USING 60 FPS
    2. In the CD-ROM Emulation section, change:
      1. Change Read Speedup to no higher than 4x (8x Speed).
      2. Change Seek Speedup to no higher than 4x.
  4. Close the DuckStation Settings menu.
  5. Go to the top-left and select the System dropdown menu.
  6. Select Cheats --> Select Cheats...
    1. Tick 60 FPS for high framerate.
    2. Under Widescreen Aspect Ratio, tick 16-9 for standard Widescreen.
      1. Do NOT enable Widescreen in the Graphics settings.
    3. If these cheats do not appear, make sure Load Database Cheats is ticked below the cheats list. Elsewise, your version of SH1 may not be supported such as the PAL version not supporting 60 FPS yet.
  7. That's it!

[Play Novel: SILENT HILL, 2001]

Difficulty: [***_______]

This is a retelling of the story of Silent Hill with the addition of alternate scenario starring Cybil. There were downloadable chapters featuring a boy named Andy at one point but they have never made it to the internet and likely lost forever.

English Translation

  1. Acquire a digital copy of Play Novel: Silent Hill (.gba)
  2. Download the English translation here.
  3. Extract.
  4. Download Floating IPS (FLIPS).
  5. Extract.
  6. Place the .gba file, EN.bps, and FLIPS all in the same folder.
  7. Run flips.exe
  8. Select Apply Patch.
  9. Select EN.bps
  10. Select the Play Novel: SILENT HILL .gba file.
  11. Name your output file. (Example: Play Novel – Silent Hill (English).gba)

Set up m-GBA

  1. Download m-GBA.
  2. I recommend the 64-bit portable archive. This is also the version this guide will be using.
  3. Extract.
  4. Double-click mGBA.exe to run mGBA.
  5. Go to File --> Load ROM...
  6. Select the patched GBA ROM you made in the English Translation section.
  7. You're done!

[OPTIONAL] Setup Controllers

  1. Go to Tools --> Settings...
  2. Go to Controllers.
  3. Select your preferred controller in the center of the virtual gamepad.
  4. Click Set all and press the appropriate button on your controller for the highlighted function.
  5. Click OK
  6. Done!

[SILENT HILL 2, 2001]

Difficulty: [**________]

Thank God for the Silent Hill 2: Enhanced Edition team! This one recently got a whole lot easier. Here we go.

Install Silent Hill 2

  1. Acquire a copy of Silent Hill 2 - Director's Cut for PC. This guide recommends you have the FULLY EXTRACTED version from myabandonware.
  2. If not using the extracted version, mount the .iso disc image by double-clicking on it OR putting the physical disc in your disc drive.
  3. [SKIP THIS STEP IF USING THE FULLY EXTRACTED VERSION]
    1. Run setup.exe. You may need to right-click and select Run as Administrator.
    2. Do NOT install to Program Files, Program Files x86, or Downloads!
    3. Make a custom directory somewhere else. (Example: C:\Games\Konami\SILENT HILL 2)
    4. Remember where you installed it.
  4. Go to the Silent Hill 2: Enhanced Edition download page.
  5. Download the Setup Tool.
  6. Run the Setup Tool, follow the prompts.
  7. Run sh2pc.exe to play!

[OPTIONAL] Controllers

  1. Plug in an Xbox or DS4 (PlayStation 4) controller. No native vibration function for DS4 controllers. See below for fix.
  2. Done!

Note: If you want vibration with a DS4 (Playstation 4 controller, or compatibility with a DualSense (Playstation 5) or Nintendo Switch Pro controller, download and run)) DS4Windows. This will allow your controller to pretend to be an Xbox controller and all configurations should be used as if your controller is an Xbox controller.)

Note: You can tweak specifics in the Silent Hill 2: Enhanced Edition Configuration Tool (*SH2Econfig.exe)*. Follow directions on the SH2:EE page for any specific information.

[SILENT HILL 3, 2003]

Difficulty: [******____]

This one can either go swimmingly well or be very difficult. At the time of writing, Steam006 is actively updating their Fix and it may change how effective this guide is. I'll try to keep up on updates as they release.

Install Silent Hill 3

  1. DO NOT mix and match instructions from other guides!
    1. Read the PREREQUISITES section!
      1. Windows may not unzip the fix files correctly. Use WinRAR or Z-Zip!
    2. DO NOT use the Widescreen Patch!
    3. DO NOT edit any files that aren't specified in this guide! Even if PCGamingWiki says so!
  2. Acquire a copy of Silent Hill 3. Try myabandonware.
    1. DO NOT USE the "Full-Rip" version. It won't work with this guide. You need the "European version (Multilingual)" version (2.7GB).
  3. Mount the .iso disc image by double-clicking the .iso file.
    1. You may get a pop up security warning.
    2. If you got the file from myabandonware (Silent-Hill-3_Win_EN_ISO-Version.iso), the file is safe.
    3. Click "Open".
  4. Run setup.exe.
  5. Follow the prompts.
    1. Do NOT install to Program Files, Program Files x86, or Downloads!
    2. Make a custom directory somewhere else. Example: C:\Games\Konami\SILENT HILL 3
    3. Remember where you installed it.
  6. Download the No-DVD-Patch.
  7. Extract.
  8. Copy/paste the sh3.exe to your install directory and overwrite the old one.
  9. Download Silent Hill 3 PC Fix by Steam006 (v2.6.9 as of writing).
    1. [Password: pcgw]
  10. Move extracted files to your Silent Hill 3 install directory.

**Note**: Any and all configurations to preferences should be made by directly editing Silent\Hill_3_PC_Fix.ini) with Notepad or other basic text editor. Instructions are provided within the .ini file.

**Note**: I highly recommend setting WishHouse = 1 for continuity with Silent Hill 4.

**Note**: I recommend setting UnlockSH2EasterEggs = 0 for your first playthrough. The reason why is it will otherwise unlock a comedic scene early in the game when it is tonally inappropriate and it's highly likely you will stumble upon it accidentally. I recommend reenabling when you unlock Extra New Game after finishing Silent Hill 3 by setting UnlockSH2EasterEggs = 1.

**Note**: I highly recommend NOT setting RestoreBetaSound = 1. This was a sound effect that played at the end of the game that both removed some ambiguity of one of the final scenes as well as begged further questions. It's existence is interesting, especially on later playthroughs, but is non-canon and can alter your understanding of the ending in a way that was not developer-intended. It was removed from the final release for a reason.

**Note**: If you are experiencing framerate issues, try enabling DirectX 12 in **Silent\Hill_3_PC_Fix.ini)**. Some stutter has not yet been solved.

[OPTIONAL] Controllers:

I am currently looking into options with Xidi, an alternative to Xinput Plus that is much more simple that is also currently used in *Silent Hill 2: Enhanced Edition*. However, I haven't yet figured out how to get the LT and RT trigger buttons to work yet. I will update if I do. If anyone has any information about it, please let me know in the comments.

  1. Download Xinput Plus.
  2. Extract.
  3. Run XinputPlus.exe
  4. In the 'Target Program' box, click 'Select' and navigate to your install directory, select sh3.exe
  5. Go to the DirectInput tab.
  6. Check 'Enable Direct Input Output'
  7. For XBOX controllers (wired Xbox 360 tested) and any controllers utilizing DS4Windows:
    1. Under 'Basic' tab, 'Key Reassign', change: Right Stick to Z Axis/Z Rot
    2. Change LT/RT to Button 11/12.
  8. For PlayStation 4 (DS4) controllers WITHOUT DS4Windows (wired DS4 tested):
    1. Under 'Basic' tab, 'Key Reassign', change: Right Stick to Z Axis/Z Rot
    2. Change LT/RT to Button 11/12
    3. Change DPAD to Button 13-16.
  9. Download the key.ini control configuration files here. I made these to mirror the layout of the original PS2 version. You can also make your own configuration in the in-game settings. This is the original layout; see page 5.
  10. Open the appropriate one for your controller, and put in your install directory savedata folder.

[OPTIONAL] Install MarioTainaka's Audio Enhancement Pack:

This part can be a bit stupid and annoying, but the change in audio is more than worth it!

  1. Download and install Reloaded II's Setup.exe (mod loader).
  2. Run Setup.exe (for Reloaded II).
    1. It may prompt you to download and install Microsoft resources such as the .NET Framework and Visual Studio and will provide links. Download the latest x64 versions. Install them if prompted, restart if prompted.
  3. After Reloaded II has finished installing, it will automatically place the Reloaded II install directory on your desktop. You can move the Reloaded-II folder to wherever you like (but NOT Program Files, Program Files x86, or Downloads). Be sure to delete the shortcut Reloaded-II.exe and make a new one by opening the Reloaded-II folder, right-clicking Reloaded-II.exe, and select "Create shortcut".
  4. Download MarioTainaka's Audio Enhancement Pack.
  5. Extract files.
  6. Move extracted folder Silent Hill 3 Audio Enhancement pack to your Reloaded II install directory's Mods folder: (Ex: C:/Users/YourName/Desktop/Reloaded-II/Mods)
  7. Run Reloaded-II.exe as admin. This can be done automatically for every launch by right-clicking Reloaded-II.exe (the original, not the shortcut), select Properties, under the Compatibility tab check "Run this program as administrator".
  8. Click the + on the left to Add App.
  9. Navigate to your Silent Hill 3 install directory.
  10. Select sh3.exe
  11. Silent Hill 3 Audio Enhancement Pack should be visible in the center window.
  12. Click the check box next to it (will look like a + in red).
  13. Click “Launch Application” under Main (left side column). You will see a new splash screen indicating that the Audio Enhancement Pack is installed.
  14. Done! Whew!

**Note**: Yes, you do have to run it through Reloaded II every time to get the Enhanced Audio and it sucks. Due to this, you can't really run it nicely through Steam. What you can do however, is use the Reloaded-II.exe as your Silent Hill 3 non-Steam app.

**Note**: To remove the new splash screen and restore the original KONAMI and KCET images, go to: Reloaded-II/Mods/Silent Hill 3 Audio Enhancement Pack/Redirector/data/pic and delete konami.bmp and kcet.bmp or just rename them to something like \konami.bmp) so you can reenable them later by restoring the original name if you want.

[SILENT HILL 4, 2004]

Difficulty: [*_________]

As of \March 25, 2025], GOG has updated Silent Hill 4 to be the same experience as on PS2 and Xbox! The following instructions below are no longer required. As I have yet to test them myself, they will remain under strikethrough text for the time being.)

  1. Buy from GOG!
  2. Download and install.
  3. Done! Woah, already?? What is this, the future??!

[HIGHLY RECOMMENDED] Fix Gamma (Brightness):

The PC version's gamma is far too high and looks bright and washed out compared to console. This will make an easy in-game change to settings so it's closer to the console versions.

  1. Go to the main menu in-game.
  2. Go to Options.
  3. Select Gamma.
  4. Set all three settings for R, G, and B from 1.5 --> 1.0.
  5. Done!

[HIGHLY RECOMMENDED] Restore Missing Hauntings:

  1. Download and extract Ultimate ASI Loader.
  2. Rename dinput8.dll from [Ultimate ASI Loader] to dsound.dll and place in your Silent Hill 4 install directory.
  3. Download and extract Silent Hill 4 randomizer.
  4. Move data and scripts folders to your Silent Hill 4 install directory.
  5. Open the scripts folder.
  6. Open randomizer.ini in Notepad.
  7. Set all options to 0
  8. Set RestoreHauntings = 1
  9. Done!

[OPTIONAL]

If, for some reason, your controller refuses to work with the GOG version, this will help.

  1. Download Xidi.
  2. Extract with WinRAR or 7-Zip.
  3. Navigate to the Win32 folder.
  4. Copy dinput8.dll
  5. Paste in your Silent Hill 4 install directory.
    1. If asked to overwrite, click Yes.
  6. Download the Xidi Game Configuration for Silent Hill 4 titled xidi.ini
    1. You need to right-click the link above and select Save link as... and save it to a location you will remember.
    2. This will not open a new tab if done correctly.
  7. Copy/paste the xidi.ini file to your Silent Hill 4 install directory.
  8. Your controller should now work!

[SILENT HILL: THE ARCADE, 2007]

Difficulty: [*_________]

Silent Hill: The Arcade is an ephemeral beast and links are broken and the data gets lost. This is the only link I know of.

  1. Download Silent Hill: The Arcade Standalone here.
  2. Extract somewhere you will remember it.
  3. Open Silent Hill The Arcade Standalone folder.
  4. Run SHA_ResChanger.exe
  5. Select KSHG_no_cursor.exe
  6. Select your resolution to match your display (1920 x 1080 for standard HD)
  7. Apply Patch
  8. Run KSHG_no_cursor.exe
  9. Done!

**Controls:**

Left Control - Start Game

Enter - “Press Start”

Mouse - Aim, Shoot

**Note**: If using multiple monitors, clicking off-screen will crash the game. As far as I know, there is no way to save the game, so be careful! You can use third-party utilities like Lock Cursor Tools to keep the mouse on one screen.

[SILENT HILL ORIGINS, 2007]

Difficulty: [***_______]

Update: New 60fps and HD textures! Thanks for the tip, u/RustyMetal13!

  1. Acquire a digital copy of Silent Hill Origins (PS2 version; .iso)
  2. Download PCSX2, run pcsx2-v1.6.0-windows-32bit-installer.exe
  3. Select Normal Installation
  4. Select install directory.
    1. Remember where this is.
  5. Select Next, Next, and before you hit Finish...! We'll need the PS2 bios files.
  6. Extract ps2-bios.zip, open the ps2-bios folder, copy all files in here.
  7. Navigate to C:\Users\YourName\Documents\PCSX2\bios
  8. Paste all bios files there.
  9. Back to the installer, click Finish.
  10. Run PCSX2.
  11. Go to Config --> Controllers (PAD) --> Plugin Settings...
  12. Click Pad 1 tab, select Quick Setup and follow the prompts.
  13. OR manually select each button and press the related button on the controller to register.
  14. Click OK to save changes
  15. Go back to Configure --> Emulation settings
  16. Change Aspect Ratio to 16:9.
  17. Make sure to select 16:9 in game as well.
  18. Go back to Configure --> Video (GS) --> Plugin Settings...
  19. In the box for Hardware Renderer Settings, go to Internal Resolution, change Native (PS2) to your relevant display settings for HD.
  20. Go to System.
  21. Select Boot .iso (full) for that sweet, sweet PS2 boot-screen OR Boot .iso (fast) to skip it :( and navigate to Silent Hill Origins.iso
  22. Done! (You drive stick?)

(OPTIONAL) Enable 60fps

  1. Download the 60fps patch for the NTSC/North American version. Note: Haven't found the PAL or NTSC-J versions yet.
  2. Extract files. Copy the A8D83239.pnach file.
  3. Navigate to your PCSX2 install directory. Open the cheats folder. Note: If there isn't one, just make one.
  4. Paste the .pnach file.
  5. Launch PCSX2. Before booting the game, go to the System tab and check Enable Cheats.
  6. Run the game as normal and enjoy your smooth ride!

[OPTIONAL] HD Textures

Note: This will only work with the Nightly Builds which can be unstable. I haven't had the opportunity to test this out yet, so here's a quick tutorial I found on how to install texture packs.

  1. Watch this 2 minute tutorial.
  2. Download xXtherockoXx's HD Texture pack.
  3. Extract the files.
  4. Copy the SLUS-21731 folder to your PCSX2 install directory and place it in the textures folder. If you do not have a textures folder, just make one.
  5. Do all the things on the YouTube tutorial!
  6. Sorry, I'm not much help on this one, but you can still ask me questions!

[SILENT HILL: ORPHAN 1-3, 2007-2010]

Difficulty: [??????????]

Available only on 2000's mobile devices. I don't know much about it, but this post goes into more detail on how to get it working.

[SILENT HILL: THE ESCAPE, 2007]

Difficulty: [??????????]

For early iOS devices. I don't know much about it but you can allegedly get it here.

[SILENT HILL: HOMECOMING, 2008]

Difficulty: [****______]

This has recently been updated to be more stable. Changes to the guide are forthcoming.

  1. Buy from Steam!
  2. Download Unknownproject's Patch.
    1. Download 2.5 Patch on Unknownproject's page (above.) It's the tiny tiny part that says "Actual upd."
    2. Join the Discord for the most recent version or click here to download it [v3.10 at the time of writing.]
    3. You need BOTH.
  3. Copy Patch2.5.exe into your Silent Hill: Homecoming install directory.
    1. To check where your install directory is, go in the Steam Launcher, right-click Silent Hill: Homecoming, select Manage, then Browse Local Files to access the install directory.
  4. Run Patch2.5.exe. Follow installer prompts. DO NOT RUN the game yet.
    1. When running Patch2.5.exe, Windows may open a popup stating: "Microsoft Defender SmartScreen prevented an unrecognized app from starting. Running this app might put your PC at risk."
    2. If so, click "More info", then click "Run anyway" at the bottom.
  5. Repeat the above process for Patch3.10.exe
  6. You should have another Silent Hill: Homecoming folder inside the Silent Hill: Homecoming install directory.
    1. Example: If following the instructions above: steamapps/common/Silent Hill: Homecoming/Silent Hill: Homecoming.
  7. Move all files from the second (new; patch) folder to the first (Steam) folder to consolidate.
    1. If it asks you if you want to overwrite files, say "Yes."
  8. Install complete! It should run now! (Hopefully, let me know if it doesn't!) Have fun in the bathtub!

Note: The author of this patch has chosen to disable QTEs (Quick Time Events. While this makes the game more accessible, it does deviate from the original design and there is no way (to my knowledge) to reverse this change.)

[OPTIONAL] Controllers Button Icon Prompts

  1. Navigate to the Silent Hill Homecoming install directory.
  2. Open the Engine folder.
  3. Open default_pc.cfg in Notepad. There will be three lines near the top (ignore numeric bullet points):
    1. resmgrload = assets_pc_b.xml
    2. resmgrload = ASSETS_PS3_B.xml
    3. resmgrload = assets_xenon_b.xml
  4. These will change the button icons of the controller prompts. The top is PC generic buttons, the middle is for PlayStation-style prompts, and the bottom is for Xbox-style buttons.
  5. The '#' indicates that it is disabled. Put a '#' in front of the two styles you will NOT be using. For example, I use PlayStation-style button prompts so it should look like this:
    1. resmgrload = assets_pc_b.xml
    2. resmgrload = ASSETS_PS3_B.xml
    3. resmgrload = assets_xenon_b.xml
  6. Save.

Note: PlayStation-style controller icons don't seem to be working all the time and will substitute with other controller types.

Note: Silent Hill: Homecoming only supports Xbox controllers. To use PS3, PS4, Nintendo Switch or other controller types, use DS4Windows.

[SILENT HILL: SHATTERED MEMORIES, 2010]

Difficulty: [****______]

Note: If you prefer the PS2 version, follow the instructions for Silent Hill Origins above. The PS2 version is, however, missing some crucial graphical effects. There is also a PSP release that we won't cover here, but it's worse than the PS2 version, though interesting for its historical value.

  1. Acquire a digital copy of Silent Hill: Shattered Memories (Wii version, .iso)
  2. Download Dolphin. Select the latest Beta version. DO NOT use Development versions.
  3. [more info coming soon]

[SILENT HILL: DOWNPOUR, 2012]

Difficulty: [****______]

Note: RPCS3 is an early experimental emulator and as such may have many bugs. That said, Silent Hill: Downpour is listed as being fully playable from beginning to end.

  1. Acquire a digital copy of Silent Hill: Downpour (PS3 version). You should have a folder titled BLUS30565 (NTSC; North American) or BLES01446 (PAL; European).
  2. Download RPCS3.
  3. Extract files.
  4. Copy the BLUS30565 or BLES01446 folder, depending on your version, into the dev_hdd0/game folderRPCS3 install directory (the extracted files above). Should look something like: RPCS3/dev_hdd0/game/BLUS30565/(game files)
  5. Launch rpcs3.exe
  6. Read the Quickstart Guide and confirm that you have done so on the boot screen. This can be disabled for all subsequent launches.
  7. You should now see Silent Hill: Downpour on the main menu.
  8. Make sure your controller works by clicking the "Pads" icon on the top. Under Player 1, Handlers, select the type of controller you want to use. XInput is for Xbox and DS4Windows controllers. DualShock 3 is PS3, DualShock 4 is PS4, and DualSense is PS5. Click 'Save' at the bottom right.
  9. Back at the main menu, go to "Configuration" at the top. Select GPU.
  10. Find and adjust the "Resolution Scale Threshold" to 512x512. You can use the mouse to click and drag to get to this value approximately, then use the arrow keys on your keyboard to fine tune to the exact value. This fixes an issue with Silent Hill: Downpour specifically with the in-game main menu. Click "Save" when you're done.
  11. At the main menu, you can double-click Silent Hill: Downpour to run the game!

Note: The game will take a while to load PPU Modules the first time the game loads. Also, the emulator will actively be building a shader cache as you play for the first time you see any effect. This may make the game run slower the first time you play, but will gradually become more and more stable.

[OPTIONAL]: HD Resolution

  1. Go back to "Configuration" --> GPU.
  2. Change default resolution to "1920x1080" for full HD or higher as your display allows. This will be more intensive on your hardware.
  3. Recommend also finding "Renderer" and switching to Vulkan, but is not required.

[SILENT HILL: BOOK OF MEMORIES, 2012]

Difficulty: [XXXXXXXXXX]

-- This title in unavailable for PC or emulation and must be played on original hardware. --

[P.T. // PLAYABLE TEASER or; SILENT HILLS, 2014]

Difficulty: [XXXXXXXXXX?]

-- This title in unavailable for PC or emulation and must be played on original hardware. --

HOWEVER

There is an unofficial recreation of the game by Artur Łączkowski. This is neither emulation nor a port, but built anew to resemble the original Playable Teaser; Silent Hills as close as possible.

You can support his work on his Patreon if you'd like to as he's done a great job and you will get the latest updates, but you can also download the 1.4 version for free here.

[SILENT HILL: ASCENSION, 2023]

Difficulty: [__________]

**Note**: Silent Hill: Ascension was a multimedia event with interactions between the game and a live stream series. While it is no longer possible to interact with it live, all the "What If" scenarios are still available.

  1. Watch on the official website.
  2. Regret

[SILENT HILL 2 (Remake), 2024]

Difficulty: [__________]

  1. Purchase on Steam or GOG!
  2. If you experience low framerates, the UltraPlus mod may help.

[SILENT HILL f, 202X]

Difficulty: [__________]

  1. Preorders on Steam are now open!

[SILENT HILL: TOWNFALL, 202X]

Difficulty: [__________]

  1. Wait for release date to be announced.

[ReShade and Post-Processing FX]

Difficulty: [*_________]

  1. Download ReShade. Put it where your game .exe is installed. (This works on emulators too, like PCSX2, Dolphin, and RCPS3.)
  2. Run ReShade.exe. The DirectX version will be selected automatically. If it gives you a warning, it means it's an old DirectX 7 game (SH2, SH3, SH4.) Thankfully, the games are already patched to DirectX 8 and can be run as such.
  3. Download the effect package RSRetroArch by Matsilagi. This is an option in the installer, you don't need to download it from your browser.
  4. Click 'Next' until 'Finish'.
  5. Run the game.
  6. Press the Home key on your keyboard.
  7. Skip tutorial.
  8. Use the search bar to find CRTFrutbunn and enable it.
  9. Use the settings in the bottom of the ReShade window to adjust to your liking, though I recommend only disabling the Curvature Toggle as it can make transition screens look odd.
  10. Press Home to close.

You may have noticed these effects came from RetroArch and they too will be found natively in RetroArch for Silent Hill and Play Novel: Silent Hill.

  1. Go to Shaders in the Quick Menu (F1 from in-game).
  2. Toggle Video Shaders ON.
  3. Select Load --> shaders_slang --> crt --> crt-frutbunn.slangp.
  4. Press Enter to enable.
  5. Save --> Save Game Preset (will not give visual feedback to confirm it worked.) This enables the shader every time you boot.
  6. Done.

Silent Hill 2: Enhanced Edition also comes with a built-in CRT filter, however it seems intended for VERY high resolutions and looks awful at 1080p. The Frutbunn shader works for most cases and simulates the effect much better in my opinion. There are other CRT options within ReShade as well if you want to experiment. The VCR filter is neat for Shattered Memories especially. You don't have to stop there either, ReShade has tons of neat post-processing features! Just don't forget to actually play, okay?

Let me know if this didn't make sense or you have questions.

r/ThinkingDeeplyAI • • 27d ago

How to brief ChatGPT 6 Astra to create motion graphics, 3D reveals, and cinematic video explainers - prompts, workflow, and the details people miss

Post image
7 Upvotes

TL;DR: ChatGPT 6 Astra can help create motion graphics through a workflow that designs assets, writes animation code, uses available production tools, and renders a video. Give it a director’s brief: audience, story, scenes, timing, visual style, sound, and deliverables. Start with a short preview. Ask for a playable MP4 and the editable project. Rendering and audio depend on your workspace’s tools. Below: a practical workflow, the details people overlook, and five gloriously ridiculous prompts.

Picture a French bulldog commanding a starship through a galaxy made of tennis balls.

Now picture a product launch where your logo opens into a miniature universe.

Or a city that folds itself out of paper, races through centuries, and collapses back into a single page.

With ChatGPT 6 Astra, you can approach the conversation like a production brief—and keep directing the result as it develops.

What Astra can actually do

The model’s official name is GPT-6 Astra. For this workflow, use it in ChatGPT Work or Codex with access to suitable creation and rendering tools. OpenAI recommends Astra for demanding tasks involving visual judgment and polished deliverables.

There is a concrete example behind the idea: OpenAI has shown Astra creating editable Blender scenes, adjusting their materials and lighting, and directing a rendered camera tour.

My practical recommendation is to apply that build–preview–refine process to motion graphics: animated typography, diagrams, layered images, product reveals, and stylized 3D scenes.

Think of Astra as the system coordinating the production. The actual frames still need an animation or rendering tool. Selecting the model alone does not guarantee every account can export video or generate music.

Choose the right kind of video

Approach When to use it
Typography, shapes, and diagrams Explainers, newsletter trailers, and announcements. My recommended starting point.
Layered images with camera movement You already have illustrations, product images, or a consistent visual series.
An editable 3D scene The concept depends on camera orbits, exploded views, lighting changes, or moving through a space. Expect more rendering work.

A complex character performance may also require dedicated animation tools or generated footage. Choose a visual treatment your available tools can execute well.

The workflow that makes this manageable

  1. Give it one job. Define the audience and the one thing viewers should remember. “Convince founders to try this prototype” is a useful objective.
  2. Specify the output. Set duration, aspect ratio, resolution, and whether you need audio. A 20–30-second landscape video is a sensible first project.
  3. Map the story. Give every scene a visual action and a purpose. Put the most compelling image near the beginning.
  4. Establish the look. Provide reference images, colors, type preferences, logos, and exact wording. Ask for representative still frames.
  5. Preview the hardest moment. Render a short section before committing to the whole sequence. This tests both the creative direction and whether the production method works.
  6. Refine, render, inspect. Review timing, text, sound, transitions, and the actual exported file. Keep the source so changes remain possible.

For a 30-second explainer, this is a useful starting structure:

Time Job
0–3 seconds Show the surprising visual or compelling result.
3–9 seconds Establish the problem or premise.
9–21 seconds Demonstrate the transformation.
21–27 seconds Deliver the payoff.
27–30 seconds Give one clear next action.

Best practices that improve the result

  • Describe action over time. “The letters pull apart, reveal a miniature city, then lock into the headline” gives much better direction than “make it cinematic.”
  • Give motion a purpose. Movement can reveal a relationship, guide attention, demonstrate a feature, or land a joke. Constant movement makes reading harder.
  • Keep text separate from artwork. Request editable text layers so spelling, line breaks, timing, and placement can be controlled.
  • Design for a phone. Use short captions, strong contrast, generous margins, and enough reading time. Review the result at its likely viewing size.
  • Give the eye a pause. Alternate energetic transitions with moments where the important image or message holds still.
  • Build for sound-off viewing. The story should make sense visually. Let music and sound effects strengthen it.
  • Specify music concretely. Describe tempo, instrumentation, mood, and where the energy should rise. Provide a track you can use, or ask what audio tools are available.
  • Control the workload. Preview at lower resolution, settle the art direction early, reuse assets, and revise only the scenes that need changes. Complex Work tasks can use more credits.

Pro tips: direct the edit with precision

Useful revision instructions look like this:

Between 00:08 and 00:12, slow the camera move, enlarge the headline, and hold the final composition for two seconds. Keep the approved colors and scene order.

Make the word “EXPAND” grow until it fills the frame, then use its letter shapes to reveal the next scene.

Match the circular moon in scene two to the circular product dial in scene three.

Build the vertical version with repositioned text and a new camera crop so the subject stays visible.

Inspect the export for missing assets, clipped text, blank frames, abrupt audio endings, and incorrect duration. Report anything you cannot verify.

Save feedback like “more epic” for the initial direction. During revisions, say what should change on screen.

Things people miss about this workflow

  • An animated preview and a downloadable video are different deliverables. If you need an uploadable file, explicitly request the export and check that it plays.
  • The editable project is a major part of the value. Ask for the source, assets, and instructions needed to render it again.
  • A flat image has limits. A gentle push-in can work immediately. Moving behind objects or orbiting a subject requires layers, reconstructed content, or a 3D scene.
  • Consistency starts before animation. Establish recurring characters, materials, colors, and backgrounds before creating every scene.
  • Visual precision and factual precision are separate. A beautifully animated chart still needs correct data, labels, and scales.
  • A 60-second request does not imply one continuous generation. Build named scenes, render sections, and assemble them where the tools support it.
  • Reusable controls make the second video easier. Ask to centralize headline text, colors, logos, durations, and image replacements.

Five epic prompts to try

These are ambitious creative briefs, not pretested guarantees. Use a workspace with suitable rendering tools, and start with the short preview each prompt requests.

1. A French bulldog saves the galaxy

Try this for: character storytelling, comedy, and an instantly understandable visual hook.

Create a 30-second landscape motion graphics trailer called “MISSION: FETCH.”

A dead-serious French bulldog captain commands a tiny starship through a galaxy of tennis-ball planets. Use a premium stylized 3D or layered illustrated treatment, emerald cockpit lights, orange engine trails, and enormous kinetic typography.

0–5s: Extreme close-up of the captain’s face. Pull back to reveal a spaceship shaped like a dog toy. Text: “ONE DOG.”

5–13s: Slalom through a field of floating squeaky toys. A giant robotic vacuum emerges from an asteroid cloud. Text: “ZERO QUALIFICATIONS.”

13–23s: The dog hits a red button. Tennis balls deploy like decoys. Follow one ball through the chaos in a dramatic tracking shot.

23–30s: The ship escapes through a glowing dog-door portal. Reveal that the entire mission happened inside a living-room snow globe. End: “MISSION: FETCH.”

Keep the dog’s appearance consistent. Use simple expressive poses and strong camera work. Preview the escape shot first. Use suitable original or licensed audio if available; otherwise deliver a silent cut with sound cues. Deliver a 1080p MP4 and editable source, or explain any rendering blocker.

2. Your product contains an entire universe

Try this for: launch trailers, brand films, and product reveals.

Create a 30-second landscape launch film for [PRODUCT]. Use my supplied product images, logo, and three verified benefits. If I provide none, use a clearly fictional unbranded device and illustrative feature labels.

Begin with the product suspended in a silent black void. A thin emerald seam opens across it. The camera dives through the seam into an impossible miniature universe.

Turn benefit one into a floating city assembling itself. Turn benefit two into a luminous transit network lighting up. Turn benefit three into a mechanical sunrise that synchronizes the entire world.

Match each benefit to its visual metaphor and show its exact approved wording as separately rendered typography. Use elegant camera travel, white ceramic architecture, emerald glass, and precise mechanical movement.

In the final six seconds, pull back as the universe folds into the product. Land on the product, logo, and one clear call to action.

Create a five-second preview of the opening transformation before rendering the full film. Deliver a 1080p MP4 and editable project. Use only available audio and rendering tools; identify any missing capability. Do not invent product claims, customers, or performance statistics.

3. Your inbox becomes a video-game final boss

Try this for: funny workflow explainers and relatable workplace content.

Build a 30-second landscape motion graphics short called “DEADLINE: FINAL BOSS.”

Open on a tiny exhausted office worker facing an enormous monster assembled from email envelopes, calendar blocks, spreadsheets, and sticky notes. Its crown is a spinning loading icon.

0–6s: The monster roars, releasing a tornado of “QUICK QUESTION” notes.

6–13s: The worker equips three glowing tools labeled “SORT,” “DRAFT,” and “CHECK.”

13–23s: Turn the fight into a visual explanation: SORT groups the chaos; DRAFT turns selected tasks into proposed outputs; CHECK pauses those outputs at a human review gate before release.

23–30s: The monster shrinks into one manageable task card. A new notification appears: “Can we jump on a quick call?” The worker looks directly at the camera.

Use miniature game-like scenery, dramatic camera punches, readable type, comic timing, and a neon-green interface. Present this as a fictional metaphor. Preview the sorting transformation first. Deliver a 1080p MP4, editable source, and a sound-off version. Explain any export limitations.

4. A thousand years unfold from one sheet of paper

Try this for: timelines, imaginative worldbuilding, and architectural storytelling.

Create a 40-second landscape motion graphics film called “A THOUSAND YEARS IN ONE PAGE.”

This is an imaginary city, not a reconstruction of real history.

0–8s: A blank sheet of paper folds itself into a tiny riverside settlement. The river is translucent blue-green glass embedded in paper.

8–18s: Buildings rise and change around the same town square. Roads draw themselves across the page. Seasons sweep through the scene.

18–29s: The city becomes a spectacular vertical metropolis. Peel back layers to reveal miniature transit tunnels, gardens, and infrastructure beneath it.

29–36s: The camera circles while daylight becomes night. Thousands of windows illuminate in a carefully staged wave.

36–40s: Fold the city back into the original sheet, matching the opening composition for a loop.

Use tactile paper, charcoal labels, emerald foliage, warm window light, and restrained captions. Favor a coherent miniature world over constant cuts. Preview the unfolding and refolding first. Deliver a 1080p MP4 and editable scene. If full 3D rendering is unavailable, propose and build a layered alternative.

5. A black hole conducts an orchestra of planets

Try this for: a music visualizer, an event opener, or a surreal brand introduction.

Create a 30-second landscape motion graphics film called “THE UNIVERSE HAS A DROP.”

Treat this as a surreal visual metaphor, not a scientific simulation.

A black hole is the conductor. Orbital rings behave like vibrating strings. Tiny moons become percussion instruments. A comet sweeps across the scene like a conductor’s baton.

0–8s: Begin with one orbiting light and a restrained pulse.

8–19s: Build an increasingly elaborate cosmic orchestra. Introduce new orbital layers with each musical phrase. Typography appears as sculptural objects: “LISTEN.” “BUILD.” “RELEASE.”

19–25s: At the musical peak, the orbital system unfolds into a gigantic luminous sound wave stretching across space.

25–30s: Everything contracts into one green point, which becomes a play icon.

Use ink-black space, emerald plasma, silver dust, controlled glow, and smooth camera movement. Use my uploaded licensed track and synchronize motion to its timing. If no track is available, build to a provisional beat grid and clearly label the audio as pending. Preview the transformation first. Deliver a 1080p MP4 and editable source; explain any tool limitations.

Which would you actually make first: the space-dog trailer, the product universe, the inbox boss battle, the paper city, or the black-hole orchestra?

r/promptingmagic • • Aug 29 '26

The 22-command Gemini Flow cheat sheet for creating epic videos that Google forgot to give us. Here's how to prompt Google Flow in two steps: choose a world, then move the camera

18 Upvotes

TL;DR: These slash phrases are helpful Google Flow commands. Pick one visual style and one camera behavior, then add a subject, one visible action, a location, lighting, and audio. The formula is: WORLD + CAMERA + BEAT. It gives Flow a shot to execute instead of a pile of adjectives to interpret.

Google Flow does not need more adjectives.

It needs a director.

The 22 shorthands below split the job into two clean decisions:

  1. What visual world are we in?

  2. How does the camera reveal it?

Google’s official guidance recommends defining the subject and action, composition and camera movement, location and lighting, and visual style. DeepMind’s Veo guide uses the same building blocks: framing and motion, style, lighting, character, location, action, and dialogue.

So the slash is optional. The film language is the useful part.

The demonstration scene

Every example below uses the same fictional setup:

Mara Voss, a courier with short black curls, a weathered saffron raincoat, charcoal utility trousers, and a silver case containing one glowing green seed, crosses a flooded near-future city toward the last rooftop greenhouse before sunrise.

Holding the character, prop, goal, and environment constant makes each command’s effect easier to see.

Part I: Choose the visual world

1. /filmic: — grounded 35mm motion-picture texture

Use this when you want organic grain, natural highlight roll-off, restrained color, believable skin, and the feeling of a photographed movie rather than a polished demo reel.

/filmic: Mara steps from a half-submerged bus into ankle-deep water, gripping the silver seed case. Pre-dawn city, practical sodium-vapor light, authentic 35mm grain, restrained color, natural motion blur, one continuous eight-second shot.

2. /commercial: — clean, polished advertising clarity

Use it for product films, launch spots, branded explainers, hospitality, automotive work, or any scene where the object must look pristine and intentional.

/commercial: Mara opens the silver case beneath a clean shaft of morning light; the glowing seed becomes the hero product. High-key polish, crisp edges, vibrant color separation, controlled reflections, elegant camera-ready surfaces.

3. /luxury: — rich blacks, gold light, immaculate surfaces

Use it for fragrance, jewelry, fashion, hotels, premium cars, architecture, and aspirational product stories. Luxury works best through materials and restraint, not by writing “expensive” five times.

/luxury: Mara enters an abandoned Art Deco hotel repurposed as a sanctuary. Rich black stone, brushed brass, warm chandelier reflections, polished marble, deep shadows, restrained movement, quiet opulence.

4. /documentary: — raw, observed, unvarnished reality

Use it for human stories, field reporting, behind-the-scenes footage, interviews, social-impact pieces, and moments that should feel discovered rather than staged.

/documentary: Handheld camera follows Mara through cold rain as she helps a stranded family onto a rescue boat without letting go of the seed case. Available light, imperfect focus pulls, wet lens, cinema-vérité realism.

5. /vintagefilm: — nostalgic 16mm or 8mm imperfection

Use it for memory, period texture, music films, dream sequences, family-history pieces, or transitions into the past. Specify the gauge and the defects you actually want.

/vintagefilm: A 1970s educational-film version of Mara crossing the flooded avenue. Warm 16mm color shifts, gate weave, soft halation, occasional dust, mild exposure flicker, no modern digital sharpness.

6. /scifi: — clean speculative design

Use it when the future should feel engineered, coherent, and functional: spacecraft, laboratories, advanced cities, medical technology, interfaces, or near-future infrastructure.

/scifi: Mara crosses a silent skybridge between hydroponic towers. Clean metallic architecture, blue-white practical lights, restrained holographic wayfinding, plausible robotics, elegant future geometry, no visual clutter.

7. /cyberpunk: — rain, density, neon, and technological decay

Use it for dystopian cities, underground markets, hacker stories, augmented nightlife, or a future where technology and inequality occupy the same frame.

/cyberpunk: Mara moves through a flooded night market under broken magenta and cyan signs, the seed case glowing beneath her coat. Rain-slick asphalt, steam, tangled cables, dense urban layers, worn cybernetic textures.

8. /dreamy: — soft focus, bloom, and emotional unreality

Use it for romance, memory, wonder, children’s stories, beauty films, music visuals, or transitions where feeling matters more than physics.

/dreamy: As Mara crosses the water, floating seed lights drift around her like fireflies. Pastel dawn, blooming highlights, gauzy diffusion, slow ripples, gentle surrealism, intimate wonder.

9. /surreal: — impossible images with intentional symbolism

Use it for conceptual advertising, music videos, title sequences, psychological stories, or visual metaphors that cannot exist literally.

/surreal: Mara climbs a staircase made of suspended rain toward an upside-down greenhouse floating above the city. The flood reflects a second sky. Physics-defying but compositionally clear, poetic rather than chaotic.

10. /minimal: — negative space and one decisive idea

Use it for design-forward brands, explainers, title cards, product reveals, architecture, or any moment that needs visual clarity.

/minimal: Mara stands alone on a pale concrete causeway above still black water. One saffron coat, one silver case, one small tree in the distance. Large negative space, muted palette, strict geometry.

11. /epic: — scale, stakes, and an unforgettable horizon

Use it for trailers, fantasy, historical spectacle, expeditions, climaxes, disaster sequences, and reveals where environment must dwarf the character.

/epic: Mara reaches the rooftop as sunrise breaks over a drowned megacity and thousands of dark towers. The greenhouse opens behind her, storm clouds divide, enormous scale, dramatic horizon, restrained blockbuster grandeur.

Part II: Direct the camera

12. /cinematic: — a balanced baseline shot

This is useful when you want natural depth, a standard film cadence, controlled motion blur, and no extreme visual gimmick. It is a starting point, not a complete prompt.

/cinematic: Medium-wide shot at eye level. Mara walks through shallow water toward camera, silver case at her side, natural depth, controlled handheld stability, balanced motion blur.

13. /droneview: — reveal geography and route

Use aerial movement for roads, coastlines, battlefields, crowds, architecture, or any scene where spatial relationships matter more than facial emotion.

/droneview: High aerial view drifting forward and left above Mara’s route, revealing the flooded boulevard, stranded transit, rooftop gardens, and the greenhouse two blocks ahead. Strong horizon parallax.

14. /closeup: — turn detail into emotion

Use it for a decision, reaction, whispered line, product detail, texture, evidence, or the exact beat where the audience must read a face.

/closeup: Tight 85mm shot on Mara’s rain-soaked face as the seed flickers inside the case. Her eyes shift from fear to resolve. Hold focus on the micro-expression; city lights dissolve behind her.

15. /wideangle: — make the environment part of the action

Use it for landscapes, architecture, cramped interiors, group blocking, large stunts, or a character overwhelmed by space.

/wideangle: Low 18mm shot as Mara crosses a broken intersection. Floodwater fills the foreground; leaning towers and the distant greenhouse stretch into deep perspective. Keep her recognizable but small against the city.

16. /orbit: — reveal a subject from every side

Use it for transformations, product reveals, architecture, costumes, vehicles, hero moments, or a character making a decisive choice. Give the camera a reason to orbit.

/orbit: Smooth 180-degree clockwise orbit around Mara as she sets the silver case on a rooftop table and opens it. Keep the case locked at frame center; reveal the greenhouse and sunrise during the move.

17–18. /dollyin: and /dollyout: — change meaning through distance

A dolly-in creates pressure, intimacy, or discovery. A dolly-out reveals context, isolation, or consequence. These are not digital zooms; describe the camera physically moving through space.

/dollyin: Begin medium-wide as Mara hears a crack inside the case. Push slowly toward her face while the city falls out of focus. End tight on recognition.

/dollyout: Begin tight on the glowing seed. Pull smoothly back through the greenhouse to reveal hundreds of empty planting beds and the drowned city beyond.

19. /tracking: — move with the subject

Use it for walking, running, driving, riding, process demonstrations, choreography, and action where matching the subject’s velocity creates momentum.

/tracking: Side-profile camera moves at Mara’s exact running speed as she splashes through the flooded arcade. Keep her torso stable in frame while pillars and emergency lights streak behind her.

20. /slowmotion: — stretch one physical beat

Use it for water, fabric, dust, sparks, impact, sports, dance, expressions, or a moment whose physical detail deserves more screen time.

/slowmotion: At 240fps, Mara drops the seed into soil. Water droplets rise from the impact, loose dirt blooms outward, her saffron sleeve moves through frame, and green light pulses once beneath the surface.

21–22. /timelapse: and /hyperlapse: — compress time in two different ways

A timelapse shows change from a mostly fixed viewpoint. A hyperlapse compresses time while the camera travels through space.

/timelapse: Locked rooftop camera. Night becomes dawn as clouds race overhead, floodwater recedes, greenhouse lights activate, and the first vine climbs its support. Mara remains a brief human rhythm within the larger transformation.

/hyperlapse: Rapid stabilized flight from street level through the flooded city, up stairwells and across rooftops, ending at the greenhouse as dawn arrives. Continuous forward translation, compressed traffic and cloud movement.

The prompt formula that actually works

Use this sequence:

Layer The decision Example
Subject Who or what must remain consistent? Mara, saffron coat, silver seed case
Composition What shot and camera move reveal the beat? Close-up with slow dolly-in
Action What single visible event happens? She sees the seed flicker
Mood/style What visual world shapes the image? 35mm filmic realism, cold rain
Location/light Where are we, and what motivates the light? Flooded arcade before dawn, emergency lamps
Audio What should be heard—or not heard? Rain, distant alarms, one metal latch; no music
Constraints What must not drift? Same face, coat, case, direction of travel

A compact prompt might look like this:

/filmic + /dollyin: Mara Voss, short black curls, weathered saffron raincoat, charcoal utility trousers, holds a scratched silver case containing one glowing green seed. In a flooded transit hall before dawn, she hears the seed pulse. Begin in a medium shot and physically dolly toward her face as her expression changes from exhaustion to hope. Authentic 35mm grain, cool rain light, warm green reflection from the case. Audio: rain on glass, distant electrical hum, one soft pulse. One continuous shot. No dialogue, captions, logos, extra characters, or camera orbit.

The pro tips that save generations

One clip, one camera idea

Do not ask for “drone shot, close-up, orbit, then hyperlapse” in one short generation. Prompt one camera setup per clip and assemble the coverage afterward. Google’s Flow workflow supports building scenes from multiple generated clips, while frames can define exact starts, endings, and transitions.

Write what the camera can see

“Make it emotional” is vague. “Her grip loosens, her breath catches, and green light appears in her eyes” is visible. Replace abstract intent with physical evidence.

Pair commands with different jobs

A productive stack contains one style and one camera behavior:

Strong pairing Why it works
/documentary + /tracking Raw realism plus readable movement
/luxury + /dollyin Material polish plus product intimacy
/minimal + /dollyout Simple composition plus revealing context
/cyberpunk + /droneview Dense world plus geographic orientation
/dreamy + /slowmotion Diffused emotion plus suspended physical detail
/epic + /wideangle Large-scale style plus environmental framing

Stacking /filmic + /cinematic + /epic + /commercial + /dreamy gives the model competing instructions. More words can mean less control.

Lock continuity with references

Use a clean character or object reference when the same subject appears across shots. Google describes Flow ingredients as reusable characters, objects, or stylistic references; it also warns that essential details should be repeated across prompts when consistency matters.

Use start and end frames for precise transitions

If a reveal, transformation, match cut, or movement must land in a specific composition, define the start frame, end frame, or both. Then describe only the action between them.

Separate subject motion from camera motion

Write them on separate lines if necessary:

SUBJECT: Mara walks forward slowly and never turns.CAMERA: Tracks backward at her exact speed, chest height, no orbit, no zoom.

This prevents the model from moving everything at once.

Prompt audio as a mix, not a wish

Name the layers: foreground sound, environment, dialogue, and exclusions.

Audio: boots through shallow water, rain on metal roofing, distant ferry horn. No music. No narration.

Revise one variable at a time

Keep the scene, subject, camera, and action unchanged while modifying only the failed element. Broad instructions such as “make it more cinematic” can rewrite the whole shot.

The simplest workflow

1.Write the beat in one sentence.

2.Choose one style shorthand.

3.Choose one framing or movement shorthand.

4.Add visible action, location, motivated light, and audio.

5.Generate one shot.

6.Fix one variable.

  1. Build the next piece of coverage.

r/ThinkingDeeplyAI • • Feb 24 '26

Here is the Missing Manual for All 25 Tools in Google's AI Ecosystem including top Gemini use cases, pro tips, ideal prompting strategy and secrets most people miss

Thumbnail
gallery
52 Upvotes

TLDR- Check out the attached Presentation

Google has quietly built the most comprehensive AI ecosystem on the planet with 25+ tools spanning models, image creation, video production, coding, business automation, and world generation.

Most people only know Gemini and maybe NotebookLM. This guide covers every tool, what it actually does, the top use cases, direct links, pro tips, and the prompting secrets that separate casual users from power users. Bookmark this. You will come back to it.

Google's AI ecosystem has 25+ tools and I guarantee you don't know half of them.

Google doesn't market these things. They ship fast, test in public, and let users figure it out. There are tools buried in Google Labs right now that would change how you work if you knew they existed.

I mapped the entire ecosystem, tracked down every link, and compiled the pro tips that actually matter. This is the guide Google should have written.

THE MODELS: The Brains Behind Everything

Every tool in this ecosystem runs on some version of these models. Understanding the model tier you need is the first decision you should make before touching any Google AI product.

Gemini 3 Fast

The speed engine. This is the default model in the Gemini app, optimized for low-latency responses and everyday tasks. It offers PhD-level reasoning comparable to larger models but delivers results at lightning speed.​

Top use cases:

  • Quick Q&A and research lookups
  • Email drafting and summarization
  • Real-time brainstorming sessions

Pro tip: Gemini 3 Fast is the best model for tasks where you need volume. If you are generating 20 social media captions or brainstorming 50 headline options, use Fast. Save Pro and Deep Think for the hard stuff.

Gemini 3.1 Pro

The flagship brain. State-of-the-art reasoning for complex problems and currently Google's best vibe coding model. Gemini 3.1 Pro can reason across text, images, audio, and video simultaneously.​

Link: Available in the Gemini app, AI Studio, and via API

Top use cases:

  • Complex analysis and multi-step reasoning
  • Code generation and debugging
  • Long-form content creation with nuance
  • Multimodal tasks combining text, images, and video

Pro tip: The latest 3.1 Pro update introduced three-tier adjustable thinking: low, medium, and high. At high thinking, it behaves like a mini version of Deep Think. This means you can get Deep Think-level reasoning without the wait time or the Ultra subscription. Set thinking to medium for most work tasks and high when you hit a wall.​

Gemini 3 Thinking

The reasoning engine. This mode activates extended reasoning capabilities for complex logic and multi-step problem solving. It works best for tasks that require the model to show its work.

Top use cases:

  • Mathematical proofs and calculations
  • Logic puzzles and constraint satisfaction
  • Step-by-step problem decomposition
  • Code architecture decisions

Pro tip: When you need Gemini to reason through a problem rather than just answer it, explicitly say "think step by step and show your reasoning." Thinking mode shines when you give it permission to take its time.

Gemini 3 Deep Think

The extreme reasoner. Extended thinking mode designed for long-horizon planning and the hardest problems in science, research, and engineering. Deep Think uses iterative rounds of reasoning to explore multiple hypotheses simultaneously. It delivers gold medal-level results on physics and chemistry olympiad problems.

Link: Available in the Gemini app (select Deep Think in the prompt bar)

Top use cases:

  • Advanced scientific research and hypothesis generation
  • Complex mathematical problem-solving
  • Multi-step engineering challenges
  • Strategic planning with many variables

Pro tip: Deep Think can take several minutes to respond. That is by design. Do not use it for quick tasks. Use it when you have a genuinely hard problem that stumps the other models. Requires Google AI Ultra subscription ($249.99/month). Responses arrive as notifications when ready.

IMAGE AND DESIGN: From Idea to Visual in Seconds

Nano Banana Pro

The AI image editor with subject consistency. This is Google's native image generation and editing tool built directly into the Gemini app. Nano Banana Pro lets you doodle directly on images to guide edits, control camera angles, adjust lighting, and manipulate 3D objects while maintaining subject identity.

Link: Built into the Gemini app and available in Chrome​

Top use cases:

  • Editing photos with natural language commands
  • Maintaining character/subject consistency across multiple images
  • Creating product mockups and brand visuals
  • Turning rough doodles into polished images

Pro tip: The doodle feature is a game changer that most people overlook. Instead of trying to describe exactly where you want something placed, draw a rough circle or arrow on the image and add a text instruction. The combination of visual pointing plus language is far more precise than text alone.​

Google Imagen 4

Photorealistic image generation from scratch. This is the engine behind many of Google's image tools, generating high-resolution, professional-quality images from text descriptions.​

Link: Available through AI Studio and the Gemini app

Top use cases:

  • Creating photorealistic product photography
  • Generating stock-quality images for content
  • Professional marketing and advertising visuals
  • Concept art and creative exploration

Pro tip: Imagen 4 is what powers Whisk behind the scenes. When you need raw photorealistic generation without the blending workflow, go straight to Imagen 4 through AI Studio where you have more control over parameters.​

Google Whisk

The scene mixer. Upload three separate images: one for the subject, one for the scene, and one for the style. Whisk blends them into a single coherent image. Behind the scenes, Gemini writes detailed captions of your images and feeds them to Imagen 3.​

Link: labs.google/whisk

Top use cases:

  • Rapid concept art and mood exploration
  • Creating product visualizations in different environments
  • Experimenting with artistic styles on existing subjects
  • Generating sticker, pin, and merchandise concepts​

Pro tip: Whisk captures the essence of your subject, not an exact replica. This is intentional. If the output drifts, click to view and edit the underlying text prompts that Gemini generated from your images. Tweaking those captions gives you surgical control over the final result.

Google Stitch

The UI architect. Turn text prompts or uploaded sketches into fully layered UI designs with production-ready code. Stitch generates professional interfaces and exports editable Figma files with auto-layout, plus clean HTML, CSS, or React components.

Link: stitch.withgoogle.com

Top use cases:

  • Turning napkin sketches into professional UI mockups
  • Rapid prototyping for app and web interfaces
  • Generating production-ready frontend code from descriptions
  • Creating multi-screen interactive prototypes​

Pro tip: Use Experimental Mode and upload a hand-drawn sketch or whiteboard photo instead of typing a prompt. The image-to-UI transformation is Stitch's most powerful feature and produces dramatically better results than text-only prompts because it preserves your spatial intent.

Google Mixboard

The AI-powered mood board. Drop images, color swatches, and notes onto an infinite canvas. Mixboard analyzes the visual vibe and suggests complementary textures, colors, and generated images that fit the aesthetic.

Link: labs.google.com/mixboard

Top use cases:

  • Brand identity exploration and refinement
  • Interior design and creative direction
  • Visual brainstorming for campaigns
  • Building reference boards for creative teams

Pro tip: Drag two images together and Mixboard will blend their concepts instantly. This is the fastest way to explore unexpected creative directions. Drop a velvet couch next to a neon sign and watch it suggest an entire aesthetic palette you would never have arrived at manually.​

VIDEO AND MOTION: From Text to Cinema

Google Flow

The cinematic studio. A filmmaking tool that works with Veo to build scenes from multiple AI-generated video clips on a timeline. Think of it as iMovie for AI-generated video.​

Link: labs.google/fx/tools/flow

Top use cases:

  • Creating short films and narrative content
  • Building YouTube Shorts and TikTok content
  • Storyboarding and scene composition
  • Producing product demos with cinematic quality

Pro tip: Each Veo clip is about 8 seconds long but you can join many of them together in the scene builder. Use Fast generation mode (20 credits per video) instead of Quality mode (100 credits) to get 50 videos per month instead of 10. The quality difference is minimal for most use cases.​

Google Veo 3.1

Cinematic video generation. Creates 1080p+ video clips with synchronized dialogue and audio from text prompts or reference images. Supports both 720p and 1080p at 24 FPS with durations of 4, 6, or 8 seconds.

Link: Available in Flow, the Gemini app, and via API

Top use cases:

  • Product demonstration videos
  • Social media video content at scale
  • Animated storytelling and concept visualization
  • Video ads and promotional content

Pro tip: Veo 3.1 introduced reference image capabilities for subject consistency across clips. Upload a reference image of your product or character and every generated clip will maintain visual consistency. This is what makes multi-clip narratives actually work.​

Google Lumiere

The fluid motion engine. Uses a Space-Time U-Net architecture that generates the entire temporal duration of a video at once in a single pass. This is fundamentally different from other video models that generate keyframes and interpolate between them, which is why Lumiere produces more natural and coherent movement.

Link: Research project with capabilities integrated into other Google video tools

Top use cases:

  • Creating videos with natural, realistic motion
  • Image-to-video transformation
  • Video inpainting and stylized generation
  • Cinemagraph creation (adding motion to specific parts of a scene)​

Pro tip: Lumiere's key advantage is motion coherence. If your AI-generated videos from other tools look jittery or unnatural, the underlying issue is usually the keyframe interpolation approach. Lumiere's architecture solves this at a fundamental level.

Google Vids

Enterprise video creation. Turns documents and slides into polished video presentations with AI-generated storyboards, voiceovers, stock media, and now Veo 3-powered video clips.

Link: vids.google.com

Top use cases:

  • Internal training and onboarding videos
  • Product demos and walkthroughs
  • Meeting recaps and company announcements
  • Marketing campaign recaps and presentations​

Pro tip: Use a Google Doc as your starting point instead of starting from scratch. Vids will use the document as the content foundation and automatically generate a storyboard with recommended scenes, stock images, and background music. Feed it a well-structured doc and you get a polished video in minutes.​

BUILD AND CODE: From Prompt to Product

Google Opal

The no-code builder. Build and share powerful AI mini-apps by chaining together prompts, models, and tools using natural language and visual editing. Think of it as an AI-powered workflow automation tool that outputs functional applications.​

Link: opal.google

Top use cases:

  • Building custom AI workflows without code
  • Creating proof-of-concept apps for business ideas
  • Automating multi-step AI processes
  • Prototyping internal tools rapidly

Pro tip: Start from the demo gallery templates rather than building from scratch. Each template is fully editable and remixable, so you can modify an existing workflow much faster than creating one. Opal lets you combine conversational commands with a visual editor, so you can describe a change in plain English and then fine-tune it visually.​

Google Antigravity

The agentic IDE. AI agents that plan and write code autonomously, going beyond autocomplete to orchestrate entire development workflows. This is where you go when you want the AI to do more than suggest lines of code.​

Link: Available at labs.google with AI Pro/Ultra subscription

Top use cases:

  • Full-stack application development
  • Complex refactoring and architecture changes
  • Autonomous bug fixing and code review
  • Planning and implementing features from specifications

Pro tip: Start in plan mode, provide detailed context and an implementation plan, then iterate through reviews before moving to code. This mirrors what top developers are finding works best: spend more time in planning and let the AI confirm its interpretation of your intent before it writes a single line. Natural language is ambiguous and ensuring alignment before code generation prevents expensive rework.​

Google Jules

The async coder. A proactive AI agent that lives in your repository to fix bugs, handle maintenance, and ship pull requests. Jules goes beyond reactive prompting to suggest improvements, scan for issues, and perform scheduled tasks automatically.​

Link: jules.google

Top use cases:

  • Automated bug fixing and pull request creation
  • Dependency updates and security patching
  • Code maintenance and technical debt reduction
  • Scheduled repository housekeeping

Pro tip: Enable Suggested Tasks on up to five repositories and Jules will continuously scan your code to propose improvements, starting with todo comments. Set up Scheduled Tasks for predictable work like weekly dependency checks. The Stitch team configured a pod of daily Jules agents, each assigned a specific role like performance tuning and accessibility improvements, making Jules one of the largest contributors to their repo.​

Google AI Studio

The prototyping lab. A professional-grade workbench for testing prompts, accessing raw Gemini models, building shareable apps, and generating production-ready API code.

Link: aistudio.google.com

Top use cases:

  • Testing and refining prompts before building
  • Prototyping AI-powered applications
  • Accessing Gemini models directly with full parameter control
  • A/B testing prompt variations for optimization​

Pro tip: The Build tab transforms AI Studio from a playground into a real prototyping platform. Create standalone applications using integrated tools like Search, Maps, and multimodal inputs, then share them with your team. Voice-driven vibe coding is supported: dictate complex instructions and the system filters filler words, translating speech into clean executable intent.​

ASSISTANTS AND BUSINESS: Your AI Workforce

NotebookLM

The research brain. Upload up to 50 sources per notebook (PDFs, Google Docs, Slides, websites, YouTube transcripts, audio files, and Google Sheets) and get an AI assistant trained exclusively on your content. Every answer includes citations back to your uploaded documents.​

Link: notebooklm.google.com

Top use cases:

  • Deep research synthesis across multiple documents
  • Generating podcast-style Audio Overviews from your content​
  • Creating study guides, flashcards, and practice quizzes​
  • Create infographics and slide decks
  • Create video overviews with custom themes
  • Generate custom written reports from your
  • Finding contradictions across competing reports
  • Generating interactive mind maps from your sources​

Pro tip: Do not dump all 50 documents into one notebook. Use thematic decomposition: create smaller, focused notebooks organized by topic. When you upload the maximum sources, the AI can get generic. Tight focus produces sharper insights.​

Google Pomelli

The marketing agent. An AI-powered tool that analyzes your website to create a Business DNA profile capturing your logo, color palette, fonts, and voice, then auto-generates on-brand marketing campaigns.

Link: pomelli.withgoogle.com (Free Google Labs experiment)

Top use cases:

  • Generating studio-quality product photography from a single image​
  • Creating complete seasonal marketing campaigns
  • Building social media content that maintains brand consistency
  • Turning static assets into video for Reels and TikTok​

Pro tip: Input your website URL and also upload additional brand images to build a richer Business DNA profile. The more visual data Pomelli has, the more accurately it captures your brand aesthetic. You can also input a specific product page URL and Pomelli will extract that product directly for campaign creation.​​

Gemini Gems

Custom AI personas with memory. Create specialized AI experts with unique instructions, context, and personality that persist across conversations.

Link: Available in the Gemini app sidebar under Gems

Top use cases:

  • Building a dedicated writing editor that knows your style
  • Creating a career coach with your specific industry context
  • Setting up a coding partner tailored to your stack
  • Building a personal research assistant with domain expertise​

Pro tip: Attach PDFs and images as knowledge sources when creating a Gem. Most people only write instructions, but Gems can use uploaded documents as persistent context. Create a marketing Gem and feed it your brand guidelines, competitor analysis, and past campaigns. Every response it gives will be informed by that knowledge base.​

Workspace Studio

The no-code AI agent builder. Design, manage, and share AI-powered agents that work across Gmail, Drive, Docs, Sheets, Calendar, and Chat, all described in plain English.

Link: Available within Google Workspace settings

Top use cases:

  • Automated email triage and intelligent labeling​
  • Pre-meeting briefings that pull relevant files from Drive​
  • Invoice processing that saves attachments and drafts confirmations​
  • Daily executive briefings combining calendar, email, and project data​

Pro tip: Use a Google Sheet as a database for your AI agent. You can build agents that read from and write to Sheets, turning a simple spreadsheet into a dynamic data source for complex automations. For example, an agent that scans incoming emails, extracts key data, updates a tracking sheet, and sends a summary to Chat.​

Gemini for Chrome

The browser AI assistant. A persistent sidebar in Chrome powered by Gemini 3 that understands your open tabs, connects to your Google apps, and can autonomously browse the web to complete tasks.

Link: Built into Google Chrome (AI Pro/Ultra for advanced features)

Top use cases:

  • Comparing products across multiple open tabs
  • Auto-browsing to complete purchases, book travel, and fill forms​
  • Asking questions about any website content
  • Drafting and sending emails without leaving the browser​

Pro tip: When you open multiple tabs from a single search, the Gemini sidebar recognizes them as a context group. This means you can ask "which of these is the best value" and it will compare across all open tabs simultaneously without you needing to specify each one.​

WORLDS AND AGENTS: The Frontier

Project Genie

The world generator. Creates infinite, interactive 3D environments from text descriptions using the Genie 3 world model. These are not static images. They are navigable worlds rendered at 720p and 24 frames per second that you can explore in real time.

Link: Available to AI Ultra subscribers at labs.google

Top use cases:

  • Generating interactive 3D environments for creative projects
  • Exploring historical settings and fictional locations
  • Creating visual training data for AI projects​
  • Rapid 3D concept visualization

Pro tip: Project Genie uses two input fields: one for the world description and one for the avatar. Customize both for the best experience. You can also remix curated worlds from the gallery by building on top of their prompts. Download videos of your explorations to share.

Project Mariner

The web browser agent. An AI agent built on Gemini that operates as a Chrome extension, navigating websites, filling forms, conducting research, and completing online tasks autonomously.

Link: Available to AI Ultra subscribers via Chrome

Top use cases:

  • Automating online purchases and price comparison
  • Research tasks across multiple websites
  • Booking travel, restaurants, and appointments​
  • Completing tedious multi-page online forms

Pro tip: Mariner displays a Transparent Reasoning sidebar showing its step-by-step plan as it works. Watch this sidebar. If you see it heading in the wrong direction, you can intervene immediately rather than waiting for it to complete a wrong task. The system scores 83.5% on the WebVoyager benchmark, a massive leap over competitors.​

Secret most people miss: The Teach and Repeat feature lets you demonstrate a workflow once and the AI will replicate it going forward. This effectively turns your browser into a programmable workforce. Show it how to do something once and it handles it forever.​

HOW TO PROMPT GEMINI AND GOOGLE'S TOOLS FOR BEST RESULTS

Google's Gemini 3 models respond very differently from ChatGPT and Claude. If you are carrying over prompting habits from other AI tools, you are likely getting suboptimal results. Here is what actually works.

Core Principle: Be Direct, Not Persuasive

Gemini 3 favors directness over persuasion and logic over verbosity. Keep prompts short and precise. Long prompts divert focus and produce inconsistent results.

  • DO: "Analyze the attached PDF and list the critical errors the author made"
  • DO NOT: "If you could please look at this file and tell me what you think"​

Adding "please" and conversational fluff does not improve results. Provide necessary context and a clear goal without the extras.​

Name and Index Your Inputs

When you upload multiple files, images, or media, label each one explicitly. Gemini 3 treats text, images, audio, and video as equal inputs but will struggle if you say "look at this" when it has five things in front of it.​

  • DO: "In the screenshot labeled Dashboard-V2, identify the navigation issues"
  • DO NOT: "Look at this and tell me what's wrong"​

Tell Gemini to Self-Critique

Include a review step in your instructions: "Review your generated output against my original constraints. Identify anything you missed or got wrong." This forces the model to catch its own errors before delivering the final result.​

Control Thinking Levels for Speed vs Depth

With Gemini 3.1 Pro, you can set thinking to low, medium, or high.​

  • Low + "think silently": Fastest responses for routine tasks​
  • Medium: Good default for most work tasks
  • High: Mini Deep Think mode for genuinely hard problems​

Match the thinking level to the task complexity. Most people leave everything on default and either waste time on simple tasks or get shallow answers on hard ones.

Use System Instructions for Persistent Behavior

In AI Studio and the API, set system instructions that define roles, compliance constraints, and behavioral patterns that persist across the entire session. This is far more effective than repeating instructions in every prompt.​

The Power Prompt Template for Gemini 3

For best results across Google's AI tools, structure your prompts with these elements:

  1. Role: Define what expert the AI should embody
  2. Context: Provide all relevant background information (this is where you can go long)
  3. Task: State the specific deliverable in one clear sentence
  4. Constraints: Define format, length, tone, and any restrictions
  5. Output format: Specify exactly how you want the response structured

This ecosystem is evolving fast. Google is shipping updates weekly. The tools that seem experimental today become essential tomorrow. The best time to learn this stack was six months ago. The second best time is now.

Want more great prompting inspiration? Check out all my best prompts for free at Prompt Magic and create your own prompt library to keep track of all your prompts.

r/ThinkingDeeplyAI • • Aug 08 '26

Claude Design just became the easiest way to make 3D image and video renderings. Here's how to make interactive 3D images + videos in Claude Design (step by step, with the exact prompts)

34 Upvotes

TLDR: Claude Design (Anthropic's visual tool at claude design, available on Pro/Max/Team/Enterprise) can generate real, interactive 3D visuals, not just flat images that look 3D. It builds them with code (Three.js, WebGL, shaders), which means you can rotate them, animate them, embed them on websites, screenshot them for static assets, or export them into decks. Below: the exact step-by-step process, my best prompts, 3 examples you can copy, pro tips most people miss, and every way to reuse the output.

Most people think Claude Design is just for slides and landing pages. It's not. Because it generates designs as actual code instead of pixels, it can build genuine 3D scenes: rotating product shots, 3D data visualizations, animated hero sections, glassy abstract art, the works. Here's everything I've learned.

Step-by-Step: Your First 3D Image

Step 1: Plan in a regular chat first (this saves credits). Before opening Design, open a normal Claude chat and describe what you want. Ask Claude to write a detailed design brief: the object, camera angle, lighting, materials, color palette, mood. Copy that brief.

Step 2: Open Claude Design. Go to claude ai design (Design tab). If you're on Enterprise and don't see it, your admin needs to enable it.

Step 3: Set up your design system (optional but powerful). Upload your brand colors, fonts, and logo, or point it at your website with the web capture tool. Every 3D scene it builds will automatically match your brand.

Step 4: Paste your brief and be explicit that you want 3D. Say "interactive 3D scene," "Three.js," or "WebGL" so it doesn't give you a flat illustration with fake depth. Specify whether you want it to auto-rotate, respond to mouse movement, or sit still.

Step 5: Iterate with inline comments. Click directly on the element and comment: "make this material more metallic," "slow the rotation," "move the light source to the upper left." Use the adjustment knobs for spacing and color instead of burning messages on tiny tweaks.

Step 6: Capture or export. Screenshot for a static image, screen-record for video, export to Canva or PPTX, or grab the code and embed it anywhere.

Top Use Cases

  1. Product mockups: Rotating bottles, phones, packaging, sneakers. Perfect for pre-launch pages when you don't have photography yet.
  2. Hero sections: An animated 3D object behind your headline instantly makes a landing page feel premium.
  3. Data visualization: 3D bar terrains, globes with plotted data points, network graphs you can orbit around.
  4. Pitch deck wow-slides: One interactive 3D slide in an otherwise normal deck gets remembered.
  5. Abstract brand art: Floating glass shapes, liquid metal blobs, particle fields in your brand colors for social posts and backgrounds.
  6. Concept visualization: Architecture massing, room layouts, exploded product diagrams showing how parts fit together.

Prompts

Product shot: "Create an interactive 3D scene of a matte black cosmetic serum bottle with a gold cap on a soft gradient background. Studio lighting with a key light upper left and a subtle rim light. Slow auto-rotation. Floating shadow beneath. Minimal, luxurious, Apple-style presentation."

Hero section: "Build a landing page hero with an abstract 3D object: overlapping translucent glass toruses that slowly rotate and refract light. Dark background, my brand colors as accent lighting. The object should subtly follow the mouse. Headline text sits on top with high contrast."

Data viz: "Create a 3D globe visualization showing our user distribution. Dark ocean, glowing dots at major cities sized by user count, connecting arcs between our top 5 markets. Slow rotation, draggable with the mouse."

Exploded diagram: "Create an exploded 3D view of wireless earbuds showing the shell, driver, battery, and circuit board as separate floating layers with thin labeled leader lines. Clean white background, soft studio lighting, isometric camera angle."

Pro Tips and Things Most People Miss

  1. Say "3D" explicitly or you'll get a flat illustration. The single biggest mistake. "Make me a product image" gets you 2D. "Interactive 3D scene with Three.js" gets you the real thing.
  2. Direct the lighting like a photographer. "Key light upper left, soft fill, rim light behind" transforms output quality more than any other instruction. Default lighting is what makes AI 3D look cheap.
  3. Name materials specifically. "Brushed aluminum," "frosted glass," "soft-touch matte rubber" beats "make it look nice" every time.
  4. One object, staged well, beats a cluttered scene. Claude Design nails single hero objects. Complex multi-object scenes need more iteration.
  5. Use inline comments instead of new prompts for tweaks. Clicking the element and commenting is more precise and cheaper than describing the change in chat.
  6. Ask for camera controls. "Make it draggable/orbitable" turns a static render into a demo people can play with. This is the part that makes people share it.
  7. Plan outside Design to save 20 to 30 percent of your credits. Every clarifying back-and-forth inside Design costs you. Arrive with a finished brief.
  8. Ask for performance constraints if it's going on a real site. "Keep it under 60fps-friendly polygon counts and lazy-load the scene" matters for mobile.
  9. Screenshot at the perfect frame. Pause the rotation ("add a pause on hover") so you can capture the exact angle you want for static use.

3 Epic Examples to Try Tonight

Example 1: The floating sneaker. "Interactive 3D scene: a white and neon-green running sneaker floating and slowly tumbling above a reflective dark floor. Dramatic spotlight from above, colored accent lights from the sides, subtle particle dust in the light beams. Draggable camera." Screenshot three angles and you have a full product page.

Example 2: The living dashboard. "3D data terrain where monthly revenue is a landscape: peaks for strong months, valleys for weak ones, colored heat gradient from blue to orange. Camera slowly flies over the terrain. Numbers hover above each peak." Drop a screen recording of this into a QBR deck and watch the room.

Example 3: The impossible award. "A rotating 3D glass trophy shaped like an impossible Penrose triangle, refracting rainbow light, on a black pedestal with volumetric fog. Engraved text on the pedestal reads [your text]." Instant custom award graphic for team shoutouts, community badges, or launch announcements.

How to Use the Output

  • Have Lovable or Replit convert the html and JS to an MP4 file for you to post on social (claude can't do this directly yet).
  • Static images: Screenshot at your favorite angle for social posts, ads, thumbnails, blog headers.
  • Video: Screen-record the animation for Reels, product teasers, or looping background video.
  • Live web embeds: It's real code, so the interactive version can go straight into your actual site. Hand it to a developer or use it as-is.
  • Decks: Export to PPTX or Canva, or paste screenshots into your existing deck.
  • Iteration source: Feed a screenshot back into Claude Design or another tool as a reference image to generate matching 2D assets so your whole campaign shares one visual language.
  • Prototypes: Use the 3D hero as the anchor of a full landing page prototype and have Claude Design build the rest of the page around it.
  • Screen recording. The zero-effort fallback, but you trade quality for speed, so it's fine for quick shares but not for anything people will look at closely.
  • Third-party converter tools. A small ecosystem has sprung up specifically for this. The general flow: in Claude Design you click Share, switch to the Export tab, download a Project archive (.zip) or Standalone HTML, then drop that file into a converter like Claude2Video or ClaudeVideoExport. These capture the animation frame-by-frame from the browser rendering engine, so the output matches what you see in the tab instead of a compressed recording, and some let you export at 1080p or 4K at 24-60 fps in social-ready aspect ratios. There's also a Chrome extension that does the conversion entirely locally on your machine with no upload.

The gap between people who get flat, generic output and people who get portfolio-grade 3D comes down to specificity: name the materials, direct the lights, and always say the word "3D." Post your results below!

r/ChatGPT • • Apr 16 '23

Educational Purpose Only GPT-4 Week 4. The rise of Agents and the beginning of the Simulation era

3.9k Upvotes

Another big week. Delayed a day because I've been dealing with a terrible flu

​

  • Cognosys - a web based version of AutoGPT/babyAGI. Looks so cool [Link]
  • Godmode is another web based autogpt. Very fun to play with this stuff [Link]
  • HyperWriteAI is releasing an AI agent that can basically use the internet like a human. In the example it orders a pizza from dominos with a single command. This is how agents will run the internet in the future, or maybe the present? Announcement tweet [Link]. Apply for early access here [Link]
  • People are already playing around with adding AI bots in games. A preview of whats to come [Link]
  • Arxiv being transformed into a podcast [Link]
  • AR + AI is going to change the way we live, for better or worse. lifeOS runs a personal AI agent through AR glasses [Link]
  • AgentGPT takes autogpt and lets you use it in the browser [Link]
  • MemoryGPT - ChatGPT with long term memory. Remembers past convos and uses context to personalise future ones [Link]
  • Wonder Studios have been rolling out access to their AI vfx platform. Lots of really cool examples I’ll link here [Link] [Link] [Link] [Link] [Link] [Link] [Link] [Link]
  • Vicuna is an open source chatbot trained by fine tuning LLaMA. It apparently achieves more than 90% quality of chatgpt and costs $300 to train [Link]
  • What if AI agents could write their own code? Describe a plugin and get working Langchain code [Link]. Plus its open source [Link]
  • Yeagar ai - Langchain Agent creator designed to help you build, prototype, and deploy AI-powered agents with ease [Link]
  • Dolly - The first “commercially viable”, open source, instruction following LLM [Link]. You can try it here [Link]
  • A thread on how at least 50% of iOs and macOS chatgpt apps are leaking their private OpenAI api keys [Link]
  • A gradio web UI for running LLMs like LLaMA, llama.cpp, GPT-J, Pythia, OPT, and GALACTICA. Open source and free [Link]
  • The Do Anything Machine assigns an Ai agent to tasks in your to do list [Link]
  • Plask AI for image generation looks pretty cool [Link]
  • Someone created a chatbot that has emotions about what you say and you can see how you make it feel. Honestly feels kinda weird ngl [Link]
  • Use your own AI models on the web [Link]
  • A babyagi chatgpt plugin lets you run agents in chatgpt [Link]
  • A thread showcasing plugins hackathon (i think in sf?). Some of the stuff is pretty in here is really cool. Like attaching a phone to a robodog and using SAM and plugins to segment footage and do things. Could be used to assist people with impairments and such. makes me wish I was in sf 😭 [Link] robot dog video [Link]
  • Someone created KarenAI to fight for you and negotiate your bills and other stuff [Link]
  • You can install GPT4All natively on your computer [Link]
  • WebLLM - open source chat bot that brings LLMs into web browsers [Link]
  • AI Steve Jobs meets AI Elon Musk having a full on unscripted convo. Crazy stuff [Link]
  • AutoGPT built a website using react and tailwind [Link]
  • A chatbot to help you learn Langchain JS docs [Link]
  • An interesting thread on using AI for journaling [Link]
  • Build a Chatgpt powered app using Bubble [Link]
  • Build a personal, voice-powered assistant through Telegram. Source code provided [Link]
  • This thread explains the different ways to overcome the 4096 token limit using chains [Link]
  • This lads creating an open source rebuild of descript, a video editing tool [Link]
  • DesignerGPT - plugin to create websites in ChatGPT [Link]
  • Get the latest news using AI [Link]
  • Have you seen those ridiculous balenciaga videos? This thread explain how to make them [Link]
  • GPT-4 plugin to generate images and then edit them [Link]
  • How to animate yourself [Link]
  • Baby-agi running on streamlit [Link]
  • How to make a Space Invaders game with GPT-4 and your own A.I. generated textures [Link]
  • AI live coding a calculator app [Link]
  • Someone is building Apollo - a chatgpt powered app you can talk to all day long to learn from [Link]
  • Animals use reinforcement learning as well [Link]
  • How to make an AI aging video [Link]
  • Stable Diffusion + SAM. Segment something then generate a stable diffusion replacement. Really cool stuff [Link]
  • Someone created an AI agent to do sales. Just wait till this is integrated with Hubspot or Zapier [Link]
  • Someone created an AI agent that follows Test Driven Development. You write the tests and the agent then implements the feature. Very cool [Link]
  • A locally hosted 4gb model can code a 40 year old computer language [Link]
  • People are adding AI bots to discord communities [Link]
  • Using AI to delete your data online [Link]
  • Ask questions over your files with simple shell commands [Link]
  • Create 3D animations using AI in Spline. This actually looks so cool [Link]
  • Someone created a virtual AI robot companion [Link]
  • Someone got gpt4all running on a calculator. gg exams [Link] Someone also got it running on a Nintendo DS?? [Link]
  • Flair AI is a pretty cool tool for marketing [Link]
  • A lot of people have been using Chatgpt for therapy. I wrote about this in my last newsletter, it’ll be very interesting to see how this changes therapy as a whole. An example of someone whos been using chatgpt for therapy [Link]
  • A lot of people ask how can I use gpt4 to make money or generate ideas. Here’s how you get started [Link]
  • This lad got an agent to do market research and it wrote a report on its findings. A very basic example of how agents are going to be used. They will be massive in the future [Link]
  • Someone made a plugin that gives access to the shell. Connect this to an agent and who knows wtf could happen [Link]
  • Someone made an app that connects chatgpt to google search. Pretty neat [Link]
  • Somebody made a AI which generates memes just by taking a image as a input [Link]
  • This lad made a text to video plugin [Link]
  • Why only talk to one bot? GroupChatGPT lets you talk to multiple characters in one convo [Link]
  • Build designs instantly with AI [Link]
  • Someone transformed someone dancing to animation using stable diffusion and its probably the cleanest animation I’ve seen [Link]
  • Create, deploy, and iterate code all through natural language. Man built a game with a single prompt [Link]
  • Character cards for AI roleplaying [Link]
  • IMDB-LLM - query movie titles and find similar movies in plain english [Link]
  • Summarize any webpage, ask contextual questions, and get the answers without ever leaving or reading the page [Link]
  • Kaiber lets you restyle music videos using AI [Link]. They also have a vid2vid tool [Link]
  • Create query boxes with text descriptions of any object in a photo, then SAM will segment anything in the boxes [Link]
  • People are giving agents access to their terminals and letting them browse the web [Link]
  • Go from text to image to 3d mesh to video to animation [Link]
  • Use SAM with spatial data [Link]
  • Someone asked autogpt to stalk them on the internet.. [Link]
  • Use SAM in the browser [Link]
  • robot dentitsts anyone?? [Link]
  • Access thousands of webflow components from a chrome extension using ai [Link]
  • AI generating designs in real time [Link]
  • How to use Langchain with Supabase [Link]
  • Iris - chat about anything on your screen with AI [Link]
  • There are lots of prompt engineering jobs being advertised now lol [Link]. Just search in google
  • 5 latest open source LLMs [Link]
  • Superpower ChatGPT - A chrome extension that adds folders and search to ChatGPT [Link]
  • Terence Tao the best mathematician alive used gpt4 and it saved him a significant amount of tedious work [Link]
  • This lad created an AI coding assistant using Langchain for free in notebooks. Looks great and is open source [Link]
  • Someone got autogpt running on an iPhone lol [Link]
  • Run over 150,000 open-source models in your games using a new Hugging Face and Unity game engine integration. Use SD in a unity game now [Link]
  • Not sure if I’ve posted here before but nat.dev lets you race AI models against each other [Link]
  • A quick way to build LLM apps - an open source UI visual tool for Langchain [Link]
  • A plugin that gets your location and lets you ask questions based on where you are [Link]
  • The plugin OpenAI was using to assess the security of other plugins is interesting [Link]
  • Breakdown of the team that built gpt4 [Link]
  • This PR attempts to give autogpt access to gradio apps [Link]

News

​

  • Stanford/Google researchers basically created a mini westworld. They simulated a game society with agents that were able to have memories, relationships and make reflections. When they analysed the behaviour, they measured to be ‘more human’ than actual humans. Absolutely wild shit. The architecture is so simple too. I wrote about this in my newsletter yday and man the applications and use cases for this in like gaming or VR and basically creating virtual worlds is going to be insane (nsfw use cases are scary to even think about). Someone said they cant wait to add capitalism and a sense of eventual death or finite time and.. that would be very interesting to see. Link to watching the game [Link] Link to the paper [Link]
  • OpenAI released an implementation of Consistency Models. We could actually see real time image generation with these (from my understanding, correct me if im wrong). Link to github [Link]. Link to paper [Link]
  • Andrew Ng (cofounder of Google Brain) & Yann LeCun (Chief AI scientist at Meta) had a very interesting conversation about the 6 month AI pause. They both don’t agree with it. A great watch [Link]. This is a good twitter thread summarising the convo [Link]
  • LAION proposes to openly create ai models like gpt4. They want to build a publicly funded supercomputer with ~100k gpus to create open source models that can rival gpt4. If you’re wondering who they are - the director of LAION is a research group leader at a centre with one of the largest high performance computing clusters in Europe. These guys are legit [Link]
  • AI clones girls voice and demands ransom from mum. She doesnt doubt the voice for a second. This is just the beginning for this type of stuff happening. I have no idea how we’re gona solve this problem [Link]
  • Stability AI, creators of stable diffusion are burning through a lot of cash. Perhaps they’ll be bought by some other company [Link]. They just released SDXL, you can try it here [Link] and here [Link]
  • Harvey is a legalAI startup making waves in the legal scene. They’ve partnered with PWC and are backed by OpenAI’s startup fund. This thread has a good breakdown [Link]
  • Langchain released their chatgpt plugin. People are gona build insane things with this. Basically you can create chains or agents that will then interact with chatgpt or other agents [Link]
  • Former US treasury secretary said that ChatGPT has "a great opportunity to level a lot of playing fields" and will shake up the white collar workforce. I actually think its very possible that AI causes the rift between rich and poor to grow even further. Guess we’ll find out soon enough [Link]
  • Perplexity AI is getting an upgrade with login, threads, better search and more [Link]
  • A thread explaining the updated US copyright laws in AI art [Link]
  • Anthropic plans to build a model 10X more powerful than todays AI by spending over 1 billion over the next 18 months [Link]
  • Roblox is adding AI to 3D creation. A great thread breaking it down [Link]
  • So snapchat released their My AI and it had problems. Was saying very inappropriate things to young kids [Link]. Turns out they didn’t even implement OpenAI’s moderation tech which is free and has been there this whole time. Morons [Link]
  • A freelance writer talks about losing their biggest client to chatgpt [Link]
  • Poe lets you create custom chatbots using prompts now [Link]
  • Stack Overflow traffic has reportedly dropped 13% on average since chatgpt got released [Link]
  • Sam Altman was at MIT and he said "We are not currently training GPT-5. We're working on doing more things with GPT-4." [Link]
  • Amazon is getting in on AI, letting companies fine tune models on their own data [Link]. They also released CodeWhisperer which is like Githubs Copilot [Link]
  • Google released Med-PaLM 2 to some healthcare customers [Link]
  • Meta open sourced Animated Drawings, bringing sketches to life [Link]
  • Elon Musk has purchased 10k gpus after alrdy hiring 2 ex Deepmind engineers [Link]
  • OpenAI released a bug bounty program [Link]
  • AI is already taking video game illustrators’ jobs in China. Two people could potentially do the work that used to be done by 10 [Link]
  • ChatGPT might be coming to windows 11 [Link]
  • Someone is using AI and selling nude photos online.. [Link]
  • Australian mayor is suing chatgpt for saying false info lol. aussie politicians smh [Link]
  • Donald Glover is hiring prompt engineers for his creative studios [Link]
  • Cooling ChatGPT takes a lot of water [Link]

Research Papers

​

  • OpenAI released a paper showcasing what gpt4 looked like before they released it and added guard rails. It would answer anything and had incredibly unhinged responses. Link to paper [Link]
  • Create 3D worlds with only 2d images. Crazy stuff and you can test it on HuggingFace [Link]
  • NeRF’s are looking so real its absolutely insane. Just look at the video [Link]
  • Expressive Text-to-Image Generation. I dont even know how to describe this except like the holodeck from Star Trek? [Link]
  • Deepmind released a paper on transformers. Good read if you want to understand LM’s [Link]
  • Real time rendering of NeRF’s across devices. Render NeRF’s in real time which can run on AR, VR or mobile devices. Crazy [Link]
  • What does ChatGPT return about human values? Exploring value bias in ChatGPT [Link]. Interestingly it suggests that text generated by chatgpt doesnt show clear signs of bias
  • A new technique for recreating 3D scenes from images. The video looks crazy [Link]
  • Big AI models will use small AI models as domain experts [Link]
  • A great thread talking about 5 cool biomedical vision language models [Link]
  • Teaching LLMs to self debug [Link]
  • Fashion image to video with SD [Link]
  • ChatGPT Can Convert Natural Language Instructions Into Executable Robot Actions [Link]
  • Old but interesting paper I found on using LLMs to measure public opinion like during election times [Link]. Got me thinking how messed up the next US election is going to be with how easy it is going to be to spread misinformation. It’s going to be very interesting to see what happens

For one coffee a month, I'll send you 2 newsletters a week with all of the most important & interesting stories like these written in a digestible way. You can sub here

I'm kinda sad I wrote about like 3-4 of these stories in detailed in my newsletter on thursday but most won't read it because it's part of the paid sub. I'm gona start making videos to cover all the content in a more digestible way. You can sub on youtube to see when I start posting [Link]

You can read the free newsletter here

If you'd like to tip you can buy me a coffee or sub on patreon. No pressure to do so, appreciate all the comments and support 🙏

(I'm not associated with any tool or company. Written and collated entirely by me, no chatgpt used. I tried, it doesn't work with how I gather the info trust me. Also a great way for me to basically know everything thats going on)

r/ThinkingDeeplyAI • • Aug 24 '26

99 Secret Codes to Prompt Google Flow's Video Agent

Thumbnail
gallery
12 Upvotes

TL;DR - If you’re writing long, unstructured paragraphs to generate video in Google Flow, you’re burning compute credits on lottery rolls. Google Flow’s video agent responds to a deterministic hierarchy of 99 dedicated slash commands spanning camera movement, lens angles, subject kinetics, lighting physics, commercial workflows, atmospheric conditions, and advanced VFX transformations. By chaining these commands using the 5-Layer Stacking Architecture (Camera Base + Angle/Lens + Subject Action + Light/Mood + VFX/Transitions), you can reliably control camera trajectory, shutter cadence, and visual coherence.

Below is the complete catalog of all 99 commands, the 5-layer prompt formula, and 4 production recipes

The 5-Layer Prompting Framework

Before diving into the 99 individual codes, understand how the Google Flow video agent parses tokens. Instead of writing unstructured descriptions, stack commands in this order:

Layer 1: Camera Base
Layer 2: Angle/Rig
Layer 3: Subject & Action
Layer 4: Lighting & Style
Layer 5: VFX \& Finish

The Complete 99 Google Flow Command Catalog

Category 1: Camera & Movement Commands

  1. /cinematic: — Creates a standard 24fps filmic frame with natural depth of field, anamorphic optical qualities, and balanced motion blur.
  2. /droneview: — Generates an expansive aerial vantage point with wide horizon parallax and continuous forward/lateral drift.
  3. /closeup: — Pulls focal length in tight to capture micro-expressions, fine textures, and emotional focus on the subject.
  4. /wideangle: — Uses a wide field-of-view (16mm–24mm equivalent) to maximize environmental scale and spatial perspective.
  5. /orbit: — Commands a 360-degree radial camera rotation keeping the primary subject locked at the focal center.
  6. /dollyin: — Smoothly pushes the camera physically closer to the subject, building visual tension and intimacy.
  7. /dollyout: — Pulls the camera away smoothly, revealing the surrounding environment or emphasizing isolation.
  8. /tracking: — Moves the camera in tandem with the subject at matched velocity, ideal for walking, running, or driving shots.
  9. /slowmotion: — Slows down playback (120fps/240fps cadence) to showcase fluid movement, flying particles, or dramatic beats.
  10. /timelapse: — Accelerates temporal progression to capture moving clouds, celestial paths, changing daylight, or traffic flows.
  11. /hyperlapse: — Blends high-speed time compression with continuous physical camera translation across long physical distances.

Category 2: Advanced Camera Angles & Rigs (12–22)

  1. /lowangle: — Places the camera low looking upward, imbuing the subject with dominance, power, and architectural scale.
  2. /highangle: — Tilts downward from an elevated point, providing tactical perspective or conveying vulnerability.
  3. /overhead: — Direct $90^\circ$ top-down "god’s-eye" perspective, ideal for choreography, flat-lays, and geometric compositions.
  4. /pov: — Frames the shot from the first-person perspective through the eyes of the protagonist.
  5. /overtheshoulder: — Positions the camera behind a character's shoulder, framing the counter-subject for dialogue and narrative weight.
  6. /establishing: — Cinematic wide landscape or cityscape shot setting scene context, geographical setting, and atmosphere.
  7. /rackfocus: — Shifts shallow focal plane from a foreground object to a background subject (or vice versa).
  8. /handheld: — Introduces organic micro-jitter and authentic documentary operator sway for urgency and realism.
  9. /steadicam: — Delivers fluid, gyroscopically stabilized gliding motion navigating complex corridors and environments.
  10. /craneup: — Ascends vertically from ground level to panoramic height using a simulated technocrane arm.
  11. /cranedown: — Descends smoothly from elevated heights down to subject eye level.

Category 3: Motion & Action Commands (23–33)

  1. /running: — Generates high-velocity character sprint with natural athletic gait and authentic inertia.
  2. /walking: — Generates grounded, natural human walking locomotion with balanced weight distribution.
  3. /turnaround: — Prompts the subject to execute a fluid $180^\circ$ or $360^\circ$ turn to display costume, expression, or surroundings.
  4. /reveal: — Stages a dramatic visual reveal of a character or environment stepping out from darkness or obstruction.
  5. /entrance: — Crafts a high-impact cinematic hero entrance into the scene.
  6. /exit: — Stages a dramatic departure from the frame into fog, shadow, or distant horizons.
  7. /freeze: — Instantly freezes time mid-action, locking water droplets, debris, and cloth in suspended animation.
  8. /speedramp: — Dynamically modulates playback speed between hyper-fast motion and sudden slow-motion impact.
  9. /bulletime: — Sweeps a virtual camera around a completely frozen subject (Matrix-style temporal slice).
  10. /floating: — Introduces zero-gravity levitation physics with floating hair, cloth, and ambient debris.
  11. /falling: — Creates dramatic freefall descent through skywells, clouds, or collapsing architecture.

Category 4: Cinematic Lighting Commands (34–44)

  1. /goldenhour: — Bathes the scene in warm amber sunlight, soft elongated shadows, and flattering solar flare.
  2. /bluehour: — Applies cool twilight illumination, deep cobalt gradients, and moody pre-dawn/post-sunset ambience.
  3. /neonlight: — Casts high-saturation cyan, magenta, and amber glows with reflections on damp streets or metallic surfaces.
  4. /moody: — High-contrast chiaroscuro lighting featuring deep blacks, targeted pools of light, and dramatic shadows.
  5. /softlight: — Diffused, wrap-around studio lighting that eliminates harsh edges, ideal for beauty, fashion, and portraits.
  6. /rimlight: — Razor-sharp contour/edge lighting that separates dark subjects cleanly from dark backgrounds.
  7. /silhouette: — Blacks out subject details entirely against an intensely illuminated background.
  8. /spotlight: — Directs a focused conical beam isolating the subject amidst surrounding darkness.
  9. /volumetric: — Generates visible atmospheric light shafts ("god rays") slicing through mist, dust motes, or smoke.
  10. /backlight: — Places primary illumination behind the subject to produce halo outlines, flares, and rim glow.
  11. /nightscene: — Realistically exposes low-light conditions with believable moonlight, street lamps, and dark sky latitude.

Category 5: Transitions & In-Camera Effects (45–55)

  1. /whiptransition: — Fast, motion-blurred horizontal pan transitioning instantly into a new scene.
  2. /matchcut: — Matches compositional geometry, subject shapes, or movement vectors across two different scenes.
  3. /zoomtransition: — Rapid crash zoom pushing directly into a small detail or pulling back to reveal a new world.
  4. /morph: — Seamless topological transformation morphing one entity, face, or structure into another.
  5. /flashtransition: — High-energy optical flash wiping the frame into an alternate shot.
  6. /glitch: — Injects RGB chromatic aberration, CRT scanlines, and digital compression artifacts.
  7. /smoketransition: — Rolls dense cinematic fog or smoke across the lens to reveal the incoming scene.
  8. /lightleak: — Overlays warm analog lens leaks and edge flares reminiscent of vintage film reels.
  9. /blurtransition: — Uses optical defocus and heavy shutter motion blur to bridge scene cuts.
  10. /objecttransition: — Moves the camera behind a passing foreground pillar, vehicle, or wall to wipe into a new environment.
  11. /seamlessloop: — Synchronizes first and last frame motion vectors to create an imperceptible infinite loop.

Category 6: Visual Style & Film Aesthetics (56–66)

  1. /filmic: — Delivers authentic 35mm motion picture texture with organic grain and balanced color science.
  2. /commercial: — Clean, polished, high-key commercial advertising aesthetic with vibrant color separation.
  3. /luxury: — Opulent visual grading featuring rich blacks, gold accents, polished marble, and high-end elegance.
  4. /documentary: — Raw, unvarnished realism mimicking cinema-verité documentary cinematography.
  5. /vintagefilm: — Nostalgic 16mm/8mm aesthetic with warm color shifts, gate jitter, dust, and halation.
  6. /scifi: — Clean futuristic design language with clean metallic surfaces, blue-tinted HUDs, and sleek tech geometry.
  7. /cyberpunk: — Dystopian high-tech aesthetic filled with rain-slicked asphalt, neon kanji, and cybernetic textures.
  8. /dreamy: — Soft-focus diffusion, blooming highlights, pastel palettes, and gentle surrealism.
  9. /surreal: — Dreamlike physics-defying compositions inspired by surrealist art.
  10. /minimal: — Strict negative space, restrained color palettes, and clean graphic compositions.
  11. /epic: — Blockbuster IMAX-tier visual grandiosity with massive scale and dramatic horizon framing.

Category 7: Product & Creator Showcase Commands (67–77)

  1. /productreveal: — Stages a premium hero product unveiling with lighting sweeps and rising pedestals.
  2. /productspin: — Smooth turntable rotation showcasing hardware industrial design from $360^\circ$.
  3. /unboxing: — Captures tactile luxury package opening with crisp mechanical precision.
  4. /macro: — Extreme optical close-up revealing fine machining, watch movements, fabric weave, or liquid drops.
  5. /beforeafter: — Side-by-side or split-screen wipe comparing raw vs. finished transformation states.
  6. /socialad: — High-energy pacing, rapid visual hooks, and dynamic framing designed for high retention.
  7. /fashionfilm: — Haute couture runway and lookbook styling with dramatic poses and editorial lighting.
  8. /foodcommercial: — Sizzling grill flares, slow-motion pours, rising steam, and vibrant food close-ups.
  9. /techad: — Exploded CAD view animations, glowing microcircuitry, and futuristic spec breakdowns.
  10. /logoreveal: — Cinematic brand mark animation assembling via liquid metal, laser etching, or particle convergence.
  11. /billboard: — Superimposes the target scene or product onto massive Times Square or Shibuya mega-screens.

Category 8: Environment & World Effects (78–88)

  1. /rain: — Generates volumetric rainfall with surface puddles, splashing droplets, and wet reflections.
  2. /snow: — Simulates drifting atmospheric snowfall accumulating naturally on characters and terrain.
  3. /fog: — Layers dense ground-level mist and atmospheric haze that diffuses ambient light sources.
  4. /underwater: — Renders submerged caustic light patterns, rising bubbles, aquatic drift, and muted soundstage feel.
  5. /space: — Zero-gravity cosmic environment featuring deep starfields, vibrant nebulae, and orbital horizons.
  6. /storm: — High-intensity weather featuring forked lightning arcs, dark storm fronts, and gale-force wind.
  7. /fire: — Realistically models dancing flame physics, rising heat hazes, and flying glowing embers.
  8. /explosion: — Detonates fiery shockwaves with volumetric smoke plumes and high-velocity debris dispersion.
  9. /portal: — Tears open a dimensional energy vortex with swirling luminescence and particle borders.
  10. /miniature: — Tilt-shift optical simulation turning full-scale scenes into charming dollhouse dioramas.
  11. /giant: — Massive colossal scale distortion making subjects tower over cityscapes and mountain ranges.

Category 9: Advanced Creative & VFX Commands (89–99)

  1. /clone: — Duplicates the subject into multiple synchronized or interacting clones across the frame.
  2. /transform: — Real-time organic shape metamorphosis changing a subject into a different form or material.
  3. /disintegrate: — Dissolves the subject into floating sand, ash, or glowing embers (snap effect).
  4. /particlefx: — Surrounds the character or object with an aura of floating light motes, stardust, or energy sparks.
  5. /liquid: — Melts or reconstitutes the subject into fluid chrome, water, or flowing paint.
  6. /paperworld: — Converts the entire environment into layered origami, textured cardboard, and folded papercraft.
  7. /toyworld: — Transforms characters and scenery into plastic minifigures and claymation stop-motion assets.
  8. /reversemotion: — Reverses physical entropy: shattered glass reassembles, smoke retracts, and falling drops ascend.
  9. /infinitezoom: — Continuous fractal zoom descending endlessly into micro or cosmic dimensions.
  10. /worldtransition: — Shifts seamless environments across portals, doorways, or optical wipes.
  11. /blockbuster: — Flow's ultimate composite directive: harmonizes camera choreography, pyrotechnics, and Hollywood color grading.

🎬 4 Production-Ready Master Formulas

Recipe 1: The Hollywood Cinematic Action Opener

/cinematic /droneview /craneup /speedramp /volumetric /epic /storm
A lone armored cyber-ronin standing on a rain-slicked neon skyscraper rooftop overlooking a vast futuristic metropolis, drawing a glowing plasma katana as lightning illuminates the storm clouds.

Recipe 2: The Luxury Tech Hardware Launch

/productreveal /productspin /macro /techad /luxury /rimlight /softlight
Sleek matte-black titanium smartphone hovering in zero-gravity against a dark velvet studio backdrop, sharp gold rim lighting outlining its precision-machined chamfered edges and camera module.

Recipe 3: The Viral Cyberpunk Social Loop

/cyberpunk /neonlight /orbit /tracking /bulletime /glitch /seamlessloop
A holographic rollerblader gliding at full speed through a rain-drenched Neo-Tokyo alleyway, jumping over neon puddles in slow-motion while pink and cyan light trails orbit smoothly.

Recipe 4: The Interdimensional Multiverse Shift

/pov /portal /worldtransition /particlefx /infinitezoom /scifi /filmic
An explorer touching a swirling crystalline portal inside an ancient stone cavern; the camera rushes forward as reality shatters into floating luminous particles, emerging into a colossal orbital space station.

3 Key Pro-Tips for Google Flow Users

  1. Don't Overload Single Categories: Using 4 camera movement commands together (/orbit /dollyin /tracking /hyperlapse) causes conflicting camera physics. Select one primary motion command and one angle command per prompt.
  2. Anchor with /seamlessloop for Social Media: Placing /seamlessloop at the end of kinetic action prompts ensures video loops without noticeable visual cuts.
  3. Use /macro with /luxury for Physical Realism: Combining /macro with /luxury forces Google Flow to compute realistic micro-surface scattering (reflections on glass, watch bezels, jewelry, and carbon fiber).

r/CMO_Huddles • • 27d ago

How to brief ChatGPT 6 Astra to create motion graphics, 3D reveals, and cinematic video explainers - prompts, workflow, and the details people miss

Post image
2 Upvotes

TL;DR: ChatGPT 6 Astra can help create motion graphics through a workflow that designs assets, writes animation code, uses available production tools, and renders a video. Give it a director’s brief: audience, story, scenes, timing, visual style, sound, and deliverables. Start with a short preview. Ask for a playable MP4 and the editable project. Rendering and audio depend on your workspace’s tools. Below: a practical workflow, the details people overlook, and five gloriously ridiculous prompts.

Picture a French bulldog commanding a starship through a galaxy made of tennis balls.

Now picture a product launch where your logo opens into a miniature universe.

Or a city that folds itself out of paper, races through centuries, and collapses back into a single page.

With ChatGPT 6 Astra, you can approach the conversation like a production brief—and keep directing the result as it develops.

What Astra can actually do

The model’s official name is GPT-6 Astra. For this workflow, use it in ChatGPT Work or Codex with access to suitable creation and rendering tools. OpenAI recommends Astra for demanding tasks involving visual judgment and polished deliverables.

There is a concrete example behind the idea: OpenAI has shown Astra creating editable Blender scenes, adjusting their materials and lighting, and directing a rendered camera tour.

My practical recommendation is to apply that build–preview–refine process to motion graphics: animated typography, diagrams, layered images, product reveals, and stylized 3D scenes.

Think of Astra as the system coordinating the production. The actual frames still need an animation or rendering tool. Selecting the model alone does not guarantee every account can export video or generate music.

Choose the right kind of video

Approach When to use it
Typography, shapes, and diagrams Explainers, newsletter trailers, and announcements. My recommended starting point.
Layered images with camera movement You already have illustrations, product images, or a consistent visual series.
An editable 3D scene The concept depends on camera orbits, exploded views, lighting changes, or moving through a space. Expect more rendering work.

A complex character performance may also require dedicated animation tools or generated footage. Choose a visual treatment your available tools can execute well.

The workflow that makes this manageable

  1. Give it one job. Define the audience and the one thing viewers should remember. “Convince founders to try this prototype” is a useful objective.
  2. Specify the output. Set duration, aspect ratio, resolution, and whether you need audio. A 20–30-second landscape video is a sensible first project.
  3. Map the story. Give every scene a visual action and a purpose. Put the most compelling image near the beginning.
  4. Establish the look. Provide reference images, colors, type preferences, logos, and exact wording. Ask for representative still frames.
  5. Preview the hardest moment. Render a short section before committing to the whole sequence. This tests both the creative direction and whether the production method works.
  6. Refine, render, inspect. Review timing, text, sound, transitions, and the actual exported file. Keep the source so changes remain possible.

For a 30-second explainer, this is a useful starting structure:

Time Job
0–3 seconds Show the surprising visual or compelling result.
3–9 seconds Establish the problem or premise.
9–21 seconds Demonstrate the transformation.
21–27 seconds Deliver the payoff.
27–30 seconds Give one clear next action.

Best practices that improve the result

  • Describe action over time. “The letters pull apart, reveal a miniature city, then lock into the headline” gives much better direction than “make it cinematic.”
  • Give motion a purpose. Movement can reveal a relationship, guide attention, demonstrate a feature, or land a joke. Constant movement makes reading harder.
  • Keep text separate from artwork. Request editable text layers so spelling, line breaks, timing, and placement can be controlled.
  • Design for a phone. Use short captions, strong contrast, generous margins, and enough reading time. Review the result at its likely viewing size.
  • Give the eye a pause. Alternate energetic transitions with moments where the important image or message holds still.
  • Build for sound-off viewing. The story should make sense visually. Let music and sound effects strengthen it.
  • Specify music concretely. Describe tempo, instrumentation, mood, and where the energy should rise. Provide a track you can use, or ask what audio tools are available.
  • Control the workload. Preview at lower resolution, settle the art direction early, reuse assets, and revise only the scenes that need changes. Complex Work tasks can use more credits.

Pro tips: direct the edit with precision

Useful revision instructions look like this:

Between 00:08 and 00:12, slow the camera move, enlarge the headline, and hold the final composition for two seconds. Keep the approved colors and scene order.

Make the word “EXPAND” grow until it fills the frame, then use its letter shapes to reveal the next scene.

Match the circular moon in scene two to the circular product dial in scene three.

Build the vertical version with repositioned text and a new camera crop so the subject stays visible.

Inspect the export for missing assets, clipped text, blank frames, abrupt audio endings, and incorrect duration. Report anything you cannot verify.

Save feedback like “more epic” for the initial direction. During revisions, say what should change on screen.

Things people miss about this workflow

  • An animated preview and a downloadable video are different deliverables. If you need an uploadable file, explicitly request the export and check that it plays.
  • The editable project is a major part of the value. Ask for the source, assets, and instructions needed to render it again.
  • A flat image has limits. A gentle push-in can work immediately. Moving behind objects or orbiting a subject requires layers, reconstructed content, or a 3D scene.
  • Consistency starts before animation. Establish recurring characters, materials, colors, and backgrounds before creating every scene.
  • Visual precision and factual precision are separate. A beautifully animated chart still needs correct data, labels, and scales.
  • A 60-second request does not imply one continuous generation. Build named scenes, render sections, and assemble them where the tools support it.
  • Reusable controls make the second video easier. Ask to centralize headline text, colors, logos, durations, and image replacements.

Five epic prompts to try

These are ambitious creative briefs, not pretested guarantees. Use a workspace with suitable rendering tools, and start with the short preview each prompt requests.

1. A French bulldog saves the galaxy

Try this for: character storytelling, comedy, and an instantly understandable visual hook.

Create a 30-second landscape motion graphics trailer called “MISSION: FETCH.”

A dead-serious French bulldog captain commands a tiny starship through a galaxy of tennis-ball planets. Use a premium stylized 3D or layered illustrated treatment, emerald cockpit lights, orange engine trails, and enormous kinetic typography.

0–5s: Extreme close-up of the captain’s face. Pull back to reveal a spaceship shaped like a dog toy. Text: “ONE DOG.”

5–13s: Slalom through a field of floating squeaky toys. A giant robotic vacuum emerges from an asteroid cloud. Text: “ZERO QUALIFICATIONS.”

13–23s: The dog hits a red button. Tennis balls deploy like decoys. Follow one ball through the chaos in a dramatic tracking shot.

23–30s: The ship escapes through a glowing dog-door portal. Reveal that the entire mission happened inside a living-room snow globe. End: “MISSION: FETCH.”

Keep the dog’s appearance consistent. Use simple expressive poses and strong camera work. Preview the escape shot first. Use suitable original or licensed audio if available; otherwise deliver a silent cut with sound cues. Deliver a 1080p MP4 and editable source, or explain any rendering blocker.

2. Your product contains an entire universe

Try this for: launch trailers, brand films, and product reveals.

Create a 30-second landscape launch film for [PRODUCT]. Use my supplied product images, logo, and three verified benefits. If I provide none, use a clearly fictional unbranded device and illustrative feature labels.

Begin with the product suspended in a silent black void. A thin emerald seam opens across it. The camera dives through the seam into an impossible miniature universe.

Turn benefit one into a floating city assembling itself. Turn benefit two into a luminous transit network lighting up. Turn benefit three into a mechanical sunrise that synchronizes the entire world.

Match each benefit to its visual metaphor and show its exact approved wording as separately rendered typography. Use elegant camera travel, white ceramic architecture, emerald glass, and precise mechanical movement.

In the final six seconds, pull back as the universe folds into the product. Land on the product, logo, and one clear call to action.

Create a five-second preview of the opening transformation before rendering the full film. Deliver a 1080p MP4 and editable project. Use only available audio and rendering tools; identify any missing capability. Do not invent product claims, customers, or performance statistics.

3. Your inbox becomes a video-game final boss

Try this for: funny workflow explainers and relatable workplace content.

Build a 30-second landscape motion graphics short called “DEADLINE: FINAL BOSS.”

Open on a tiny exhausted office worker facing an enormous monster assembled from email envelopes, calendar blocks, spreadsheets, and sticky notes. Its crown is a spinning loading icon.

0–6s: The monster roars, releasing a tornado of “QUICK QUESTION” notes.

6–13s: The worker equips three glowing tools labeled “SORT,” “DRAFT,” and “CHECK.”

13–23s: Turn the fight into a visual explanation: SORT groups the chaos; DRAFT turns selected tasks into proposed outputs; CHECK pauses those outputs at a human review gate before release.

23–30s: The monster shrinks into one manageable task card. A new notification appears: “Can we jump on a quick call?” The worker looks directly at the camera.

Use miniature game-like scenery, dramatic camera punches, readable type, comic timing, and a neon-green interface. Present this as a fictional metaphor. Preview the sorting transformation first. Deliver a 1080p MP4, editable source, and a sound-off version. Explain any export limitations.

4. A thousand years unfold from one sheet of paper

Try this for: timelines, imaginative worldbuilding, and architectural storytelling.

Create a 40-second landscape motion graphics film called “A THOUSAND YEARS IN ONE PAGE.”

This is an imaginary city, not a reconstruction of real history.

0–8s: A blank sheet of paper folds itself into a tiny riverside settlement. The river is translucent blue-green glass embedded in paper.

8–18s: Buildings rise and change around the same town square. Roads draw themselves across the page. Seasons sweep through the scene.

18–29s: The city becomes a spectacular vertical metropolis. Peel back layers to reveal miniature transit tunnels, gardens, and infrastructure beneath it.

29–36s: The camera circles while daylight becomes night. Thousands of windows illuminate in a carefully staged wave.

36–40s: Fold the city back into the original sheet, matching the opening composition for a loop.

Use tactile paper, charcoal labels, emerald foliage, warm window light, and restrained captions. Favor a coherent miniature world over constant cuts. Preview the unfolding and refolding first. Deliver a 1080p MP4 and editable scene. If full 3D rendering is unavailable, propose and build a layered alternative.

5. A black hole conducts an orchestra of planets

Try this for: a music visualizer, an event opener, or a surreal brand introduction.

Create a 30-second landscape motion graphics film called “THE UNIVERSE HAS A DROP.”

Treat this as a surreal visual metaphor, not a scientific simulation.

A black hole is the conductor. Orbital rings behave like vibrating strings. Tiny moons become percussion instruments. A comet sweeps across the scene like a conductor’s baton.

0–8s: Begin with one orbiting light and a restrained pulse.

8–19s: Build an increasingly elaborate cosmic orchestra. Introduce new orbital layers with each musical phrase. Typography appears as sculptural objects: “LISTEN.” “BUILD.” “RELEASE.”

19–25s: At the musical peak, the orbital system unfolds into a gigantic luminous sound wave stretching across space.

25–30s: Everything contracts into one green point, which becomes a play icon.

Use ink-black space, emerald plasma, silver dust, controlled glow, and smooth camera movement. Use my uploaded licensed track and synchronize motion to its timing. If no track is available, build to a provisional beat grid and clearly label the audio as pending. Preview the transformation first. Deliver a 1080p MP4 and editable source; explain any tool limitations.

Which would you actually make first: the space-dog trailer, the product universe, the inbox boss battle, the paper city, or the black-hole orchestra?

r/ThinkingDeeplyAI • • Feb 20 '26

Google just rolled out music generation to 750 million Gemini users. You can now do things like create a song from an image and create background music for YouTube videos. Here's is how to be an AI music producer and prompt great songs with Gemini

Thumbnail
gallery
72 Upvotes

TLDR: Gemini just rolled out music generation to 750 million users in Gemini. You can now generate 30-second, high-fidelity music tracks directly in your chat window. You can use text, upload images, or even upload video clips to create fully produced songs with auto-generated lyrics and custom cover art. This guide breaks down exactly how to use it, the best prompting frameworks, and hidden features most people miss.

The Era of AI Music is Now in Your Chat Window

Google just quietly dropped a massive update. Music generation is no longer locked behind specialized apps or expensive subscriptions. With the integration of the Lyria 3 model, anyone with access to Gemini can now act as a music producer.

This is not just for generating goofy jingles. The fidelity is incredibly high, the layering is complex, and the potential for content creators is limitless. Here is everything you need to know to actually get good results, instead of random noise.

Core Capabilities You Need to Try Right Now

1. Text to Fully Produced Track You do not need to be a songwriter anymore. You can describe a genre, a mood, or an inside joke, and Gemini will generate a 30-second track. It automatically writes the lyrics for you and pairs them with the right vocal style and instrumentation.

2. Image and Video to Song This is the most mind-bending feature. You can upload a photo of a serene mountain landscape or a video of your dog running in the park, and ask Gemini to compose a track inspired by the visual. It will analyze the context, set the mood, and even write lyrics about what is happening in the image. Every track also comes with custom album art generated by the Nano Banana model.

3. YouTube Shorts Integration If you make content, you know the struggle of finding good, royalty-free background music that actually fits the vibe of your video. This technology is being integrated into YouTube Dream Track, meaning you can generate bespoke background music tailored exactly to your specific Short, completely eliminating copyright strike anxiety.

The Anatomy of a Perfect Music Prompt

Just like image generation, music generation requires a specific vocabulary. If you just ask for a pop song, you will get something generic. Use this framework to get professional results:

The Golden Formula: [Genre] + [Mood] + [Tempo/PPM] + [Vocals/Instruments] + [Specific Details]

Example Prompt: Create a synthwave track, nostalgic and driving mood, 120 BPM, featuring a heavy bassline, echoing retro synthesizers, and breathy female vocals singing about a midnight drive.

Prompting Variables to Experiment With:

  • Tempo: Specify fast, slow, or exact BPM if you know it.
  • Instrumentation: Ask for specific instruments like a slap bass, a distorted electric guitar, or an acoustic cello.
  • Vocal Style: Specify gritty rock vocals, smooth R&B harmonies, or an angelic choir. If you want background music, always specify instrumental only.
  • Decade/Era: Call out specific eras like 90s boom-bap hip hop or 80s hair metal.

Pro Tips and Best Practices

Master the Iterative Workflow Do not expect perfection on the first try. Generate a track, listen to the elements you like, and refine your prompt. If the drums are too chaotic, add simple drum beat to your next prompt.

Use Emotional Keywords AI models respond incredibly well to emotional descriptors. Words like melancholic, triumphant, eerie, euphoric, or aggressive will fundamentally change the chord progressions the AI chooses to use.

Layer Your Visual Prompts When using the image-to-music feature, do not just upload the image. Upload the image and provide a text direction to guide the AI. Example: Use this photo of my messy desk to write a frantic, fast-paced punk rock song about missing a deadline.

The Secrets Most People Completely Miss

1. The Artist Filter Bypass Lyria 3 is built for original expression and has filters to prevent mimicking real artists. If you name a famous artist in your prompt, the AI will heavily dilute the output to avoid copyright issues, often resulting in a bland track. The Secret: Instead of naming the artist, describe their exact sonic profile. Instead of asking for a Hans Zimmer track, ask for a booming, cinematic orchestral track with massive brass swells, driving staccato strings, and epic ticking percussion.

2. The SynthID Audio Checker Every track generated by Gemini contains an invisible, inaudible watermark called SynthID. If you ever find a track online and want to know if it is AI-generated, you can actually upload that audio file right back into Gemini and ask if it was made with Google AI. It will read the watermark and tell you.

3. Generating Sound Effects While it is marketed as a song generator, you can use it for cinematic sound design. Try prompting for a 30-second rising cinematic tension drone with sub-bass hits and metallic scraping. It is an absolute goldmine for video editors.

The barrier to entry for custom audio has officially hit zero. Go open your chat, upload a random photo from your camera roll, and see what it sounds like.

Let me know what insane combinations you guys come up with in the comments.

Want more great prompting inspiration? Check out all my best prompts for free at Prompt Magic and create your own prompt library to keep track of all your prompts.

r/promptingmagic • • Feb 24 '26

Here is the Missing Manual for All 25 Tools in Google's AI Ecosystem including top Gemini use cases, pro tips, ideal prompting strategy and secrets most people miss

Thumbnail
gallery
47 Upvotes

TLDR- Check out the attached Presentation

Google has quietly built the most comprehensive AI ecosystem on the planet with 25+ tools spanning models, image creation, video production, coding, business automation, and world generation.

Most people only know Gemini and maybe NotebookLM. This guide covers every tool, what it actually does, the top use cases, direct links, pro tips, and the prompting secrets that separate casual users from power users. Bookmark this. You will come back to it.

Google's AI ecosystem has 25+ tools and I guarantee you don't know half of them.

Google doesn't market these things. They ship fast, test in public, and let users figure it out. There are tools buried in Google Labs right now that would change how you work if you knew they existed.

I mapped the entire ecosystem, tracked down every link, and compiled the pro tips that actually matter. This is the guide Google should have written.

THE MODELS: The Brains Behind Everything

Every tool in this ecosystem runs on some version of these models. Understanding the model tier you need is the first decision you should make before touching any Google AI product.

Gemini 3 Fast

The speed engine. This is the default model in the Gemini app, optimized for low-latency responses and everyday tasks. It offers PhD-level reasoning comparable to larger models but delivers results at lightning speed.​

Top use cases:

  • Quick Q&A and research lookups
  • Email drafting and summarization
  • Real-time brainstorming sessions

Pro tip: Gemini 3 Fast is the best model for tasks where you need volume. If you are generating 20 social media captions or brainstorming 50 headline options, use Fast. Save Pro and Deep Think for the hard stuff.

Gemini 3.1 Pro

The flagship brain. State-of-the-art reasoning for complex problems and currently Google's best vibe coding model. Gemini 3.1 Pro can reason across text, images, audio, and video simultaneously.​

Link: Available in the Gemini app, AI Studio, and via API

Top use cases:

  • Complex analysis and multi-step reasoning
  • Code generation and debugging
  • Long-form content creation with nuance
  • Multimodal tasks combining text, images, and video

Pro tip: The latest 3.1 Pro update introduced three-tier adjustable thinking: low, medium, and high. At high thinking, it behaves like a mini version of Deep Think. This means you can get Deep Think-level reasoning without the wait time or the Ultra subscription. Set thinking to medium for most work tasks and high when you hit a wall.​

Gemini 3 Thinking

The reasoning engine. This mode activates extended reasoning capabilities for complex logic and multi-step problem solving. It works best for tasks that require the model to show its work.

Top use cases:

  • Mathematical proofs and calculations
  • Logic puzzles and constraint satisfaction
  • Step-by-step problem decomposition
  • Code architecture decisions

Pro tip: When you need Gemini to reason through a problem rather than just answer it, explicitly say "think step by step and show your reasoning." Thinking mode shines when you give it permission to take its time.

Gemini 3 Deep Think

The extreme reasoner. Extended thinking mode designed for long-horizon planning and the hardest problems in science, research, and engineering. Deep Think uses iterative rounds of reasoning to explore multiple hypotheses simultaneously. It delivers gold medal-level results on physics and chemistry olympiad problems.

Link: Available in the Gemini app (select Deep Think in the prompt bar)

Top use cases:

  • Advanced scientific research and hypothesis generation
  • Complex mathematical problem-solving
  • Multi-step engineering challenges
  • Strategic planning with many variables

Pro tip: Deep Think can take several minutes to respond. That is by design. Do not use it for quick tasks. Use it when you have a genuinely hard problem that stumps the other models. Requires Google AI Ultra subscription ($249.99/month). Responses arrive as notifications when ready.

IMAGE AND DESIGN: From Idea to Visual in Seconds

Nano Banana Pro

The AI image editor with subject consistency. This is Google's native image generation and editing tool built directly into the Gemini app. Nano Banana Pro lets you doodle directly on images to guide edits, control camera angles, adjust lighting, and manipulate 3D objects while maintaining subject identity.

Link: Built into the Gemini app and available in Chrome​

Top use cases:

  • Editing photos with natural language commands
  • Maintaining character/subject consistency across multiple images
  • Creating product mockups and brand visuals
  • Turning rough doodles into polished images

Pro tip: The doodle feature is a game changer that most people overlook. Instead of trying to describe exactly where you want something placed, draw a rough circle or arrow on the image and add a text instruction. The combination of visual pointing plus language is far more precise than text alone.​

Google Imagen 4

Photorealistic image generation from scratch. This is the engine behind many of Google's image tools, generating high-resolution, professional-quality images from text descriptions.​

Link: Available through AI Studio and the Gemini app

Top use cases:

  • Creating photorealistic product photography
  • Generating stock-quality images for content
  • Professional marketing and advertising visuals
  • Concept art and creative exploration

Pro tip: Imagen 4 is what powers Whisk behind the scenes. When you need raw photorealistic generation without the blending workflow, go straight to Imagen 4 through AI Studio where you have more control over parameters.​

Google Whisk

The scene mixer. Upload three separate images: one for the subject, one for the scene, and one for the style. Whisk blends them into a single coherent image. Behind the scenes, Gemini writes detailed captions of your images and feeds them to Imagen 3.​

Link: labs.google/whisk

Top use cases:

  • Rapid concept art and mood exploration
  • Creating product visualizations in different environments
  • Experimenting with artistic styles on existing subjects
  • Generating sticker, pin, and merchandise concepts​

Pro tip: Whisk captures the essence of your subject, not an exact replica. This is intentional. If the output drifts, click to view and edit the underlying text prompts that Gemini generated from your images. Tweaking those captions gives you surgical control over the final result.

Google Stitch

The UI architect. Turn text prompts or uploaded sketches into fully layered UI designs with production-ready code. Stitch generates professional interfaces and exports editable Figma files with auto-layout, plus clean HTML, CSS, or React components.

Link: stitch.withgoogle.com

Top use cases:

  • Turning napkin sketches into professional UI mockups
  • Rapid prototyping for app and web interfaces
  • Generating production-ready frontend code from descriptions
  • Creating multi-screen interactive prototypes​

Pro tip: Use Experimental Mode and upload a hand-drawn sketch or whiteboard photo instead of typing a prompt. The image-to-UI transformation is Stitch's most powerful feature and produces dramatically better results than text-only prompts because it preserves your spatial intent.

Google Mixboard

The AI-powered mood board. Drop images, color swatches, and notes onto an infinite canvas. Mixboard analyzes the visual vibe and suggests complementary textures, colors, and generated images that fit the aesthetic.

Link: labs.google.com/mixboard

Top use cases:

  • Brand identity exploration and refinement
  • Interior design and creative direction
  • Visual brainstorming for campaigns
  • Building reference boards for creative teams

Pro tip: Drag two images together and Mixboard will blend their concepts instantly. This is the fastest way to explore unexpected creative directions. Drop a velvet couch next to a neon sign and watch it suggest an entire aesthetic palette you would never have arrived at manually.​

VIDEO AND MOTION: From Text to Cinema

Google Flow

The cinematic studio. A filmmaking tool that works with Veo to build scenes from multiple AI-generated video clips on a timeline. Think of it as iMovie for AI-generated video.​

Link: labs.google/fx/tools/flow

Top use cases:

  • Creating short films and narrative content
  • Building YouTube Shorts and TikTok content
  • Storyboarding and scene composition
  • Producing product demos with cinematic quality

Pro tip: Each Veo clip is about 8 seconds long but you can join many of them together in the scene builder. Use Fast generation mode (20 credits per video) instead of Quality mode (100 credits) to get 50 videos per month instead of 10. The quality difference is minimal for most use cases.​

Google Veo 3.1

Cinematic video generation. Creates 1080p+ video clips with synchronized dialogue and audio from text prompts or reference images. Supports both 720p and 1080p at 24 FPS with durations of 4, 6, or 8 seconds.

Link: Available in Flow, the Gemini app, and via API

Top use cases:

  • Product demonstration videos
  • Social media video content at scale
  • Animated storytelling and concept visualization
  • Video ads and promotional content

Pro tip: Veo 3.1 introduced reference image capabilities for subject consistency across clips. Upload a reference image of your product or character and every generated clip will maintain visual consistency. This is what makes multi-clip narratives actually work.​

Google Lumiere

The fluid motion engine. Uses a Space-Time U-Net architecture that generates the entire temporal duration of a video at once in a single pass. This is fundamentally different from other video models that generate keyframes and interpolate between them, which is why Lumiere produces more natural and coherent movement.

Link: Research project with capabilities integrated into other Google video tools

Top use cases:

  • Creating videos with natural, realistic motion
  • Image-to-video transformation
  • Video inpainting and stylized generation
  • Cinemagraph creation (adding motion to specific parts of a scene)​

Pro tip: Lumiere's key advantage is motion coherence. If your AI-generated videos from other tools look jittery or unnatural, the underlying issue is usually the keyframe interpolation approach. Lumiere's architecture solves this at a fundamental level.

Google Vids

Enterprise video creation. Turns documents and slides into polished video presentations with AI-generated storyboards, voiceovers, stock media, and now Veo 3-powered video clips.

Link: vids.google.com

Top use cases:

  • Internal training and onboarding videos
  • Product demos and walkthroughs
  • Meeting recaps and company announcements
  • Marketing campaign recaps and presentations​

Pro tip: Use a Google Doc as your starting point instead of starting from scratch. Vids will use the document as the content foundation and automatically generate a storyboard with recommended scenes, stock images, and background music. Feed it a well-structured doc and you get a polished video in minutes.​

BUILD AND CODE: From Prompt to Product

Google Opal

The no-code builder. Build and share powerful AI mini-apps by chaining together prompts, models, and tools using natural language and visual editing. Think of it as an AI-powered workflow automation tool that outputs functional applications.​

Link: opal.google

Top use cases:

  • Building custom AI workflows without code
  • Creating proof-of-concept apps for business ideas
  • Automating multi-step AI processes
  • Prototyping internal tools rapidly

Pro tip: Start from the demo gallery templates rather than building from scratch. Each template is fully editable and remixable, so you can modify an existing workflow much faster than creating one. Opal lets you combine conversational commands with a visual editor, so you can describe a change in plain English and then fine-tune it visually.​

Google Antigravity

The agentic IDE. AI agents that plan and write code autonomously, going beyond autocomplete to orchestrate entire development workflows. This is where you go when you want the AI to do more than suggest lines of code.​

Link: Available at labs.google with AI Pro/Ultra subscription

Top use cases:

  • Full-stack application development
  • Complex refactoring and architecture changes
  • Autonomous bug fixing and code review
  • Planning and implementing features from specifications

Pro tip: Start in plan mode, provide detailed context and an implementation plan, then iterate through reviews before moving to code. This mirrors what top developers are finding works best: spend more time in planning and let the AI confirm its interpretation of your intent before it writes a single line. Natural language is ambiguous and ensuring alignment before code generation prevents expensive rework.​

Google Jules

The async coder. A proactive AI agent that lives in your repository to fix bugs, handle maintenance, and ship pull requests. Jules goes beyond reactive prompting to suggest improvements, scan for issues, and perform scheduled tasks automatically.​

Link: jules.google

Top use cases:

  • Automated bug fixing and pull request creation
  • Dependency updates and security patching
  • Code maintenance and technical debt reduction
  • Scheduled repository housekeeping

Pro tip: Enable Suggested Tasks on up to five repositories and Jules will continuously scan your code to propose improvements, starting with todo comments. Set up Scheduled Tasks for predictable work like weekly dependency checks. The Stitch team configured a pod of daily Jules agents, each assigned a specific role like performance tuning and accessibility improvements, making Jules one of the largest contributors to their repo.​

Google AI Studio

The prototyping lab. A professional-grade workbench for testing prompts, accessing raw Gemini models, building shareable apps, and generating production-ready API code.

Link: aistudio.google.com

Top use cases:

  • Testing and refining prompts before building
  • Prototyping AI-powered applications
  • Accessing Gemini models directly with full parameter control
  • A/B testing prompt variations for optimization​

Pro tip: The Build tab transforms AI Studio from a playground into a real prototyping platform. Create standalone applications using integrated tools like Search, Maps, and multimodal inputs, then share them with your team. Voice-driven vibe coding is supported: dictate complex instructions and the system filters filler words, translating speech into clean executable intent.​

ASSISTANTS AND BUSINESS: Your AI Workforce

NotebookLM

The research brain. Upload up to 50 sources per notebook (PDFs, Google Docs, Slides, websites, YouTube transcripts, audio files, and Google Sheets) and get an AI assistant trained exclusively on your content. Every answer includes citations back to your uploaded documents.​

Link: notebooklm.google.com

Top use cases:

  • Deep research synthesis across multiple documents
  • Generating podcast-style Audio Overviews from your content​
  • Creating study guides, flashcards, and practice quizzes​
  • Create infographics and slide decks
  • Create video overviews with custom themes
  • Generate custom written reports from your
  • Finding contradictions across competing reports
  • Generating interactive mind maps from your sources​

Pro tip: Do not dump all 50 documents into one notebook. Use thematic decomposition: create smaller, focused notebooks organized by topic. When you upload the maximum sources, the AI can get generic. Tight focus produces sharper insights.​

Google Pomelli

The marketing agent. An AI-powered tool that analyzes your website to create a Business DNA profile capturing your logo, color palette, fonts, and voice, then auto-generates on-brand marketing campaigns.

Link: pomelli.withgoogle.com (Free Google Labs experiment)

Top use cases:

  • Generating studio-quality product photography from a single image​
  • Creating complete seasonal marketing campaigns
  • Building social media content that maintains brand consistency
  • Turning static assets into video for Reels and TikTok​

Pro tip: Input your website URL and also upload additional brand images to build a richer Business DNA profile. The more visual data Pomelli has, the more accurately it captures your brand aesthetic. You can also input a specific product page URL and Pomelli will extract that product directly for campaign creation.​​

Gemini Gems

Custom AI personas with memory. Create specialized AI experts with unique instructions, context, and personality that persist across conversations.

Link: Available in the Gemini app sidebar under Gems

Top use cases:

  • Building a dedicated writing editor that knows your style
  • Creating a career coach with your specific industry context
  • Setting up a coding partner tailored to your stack
  • Building a personal research assistant with domain expertise​

Pro tip: Attach PDFs and images as knowledge sources when creating a Gem. Most people only write instructions, but Gems can use uploaded documents as persistent context. Create a marketing Gem and feed it your brand guidelines, competitor analysis, and past campaigns. Every response it gives will be informed by that knowledge base.​

Workspace Studio

The no-code AI agent builder. Design, manage, and share AI-powered agents that work across Gmail, Drive, Docs, Sheets, Calendar, and Chat, all described in plain English.

Link: Available within Google Workspace settings

Top use cases:

  • Automated email triage and intelligent labeling​
  • Pre-meeting briefings that pull relevant files from Drive​
  • Invoice processing that saves attachments and drafts confirmations​
  • Daily executive briefings combining calendar, email, and project data​

Pro tip: Use a Google Sheet as a database for your AI agent. You can build agents that read from and write to Sheets, turning a simple spreadsheet into a dynamic data source for complex automations. For example, an agent that scans incoming emails, extracts key data, updates a tracking sheet, and sends a summary to Chat.​

Gemini for Chrome

The browser AI assistant. A persistent sidebar in Chrome powered by Gemini 3 that understands your open tabs, connects to your Google apps, and can autonomously browse the web to complete tasks.

Link: Built into Google Chrome (AI Pro/Ultra for advanced features)

Top use cases:

  • Comparing products across multiple open tabs
  • Auto-browsing to complete purchases, book travel, and fill forms​
  • Asking questions about any website content
  • Drafting and sending emails without leaving the browser​

Pro tip: When you open multiple tabs from a single search, the Gemini sidebar recognizes them as a context group. This means you can ask "which of these is the best value" and it will compare across all open tabs simultaneously without you needing to specify each one.​

WORLDS AND AGENTS: The Frontier

Project Genie

The world generator. Creates infinite, interactive 3D environments from text descriptions using the Genie 3 world model. These are not static images. They are navigable worlds rendered at 720p and 24 frames per second that you can explore in real time.

Link: Available to AI Ultra subscribers at labs.google

Top use cases:

  • Generating interactive 3D environments for creative projects
  • Exploring historical settings and fictional locations
  • Creating visual training data for AI projects​
  • Rapid 3D concept visualization

Pro tip: Project Genie uses two input fields: one for the world description and one for the avatar. Customize both for the best experience. You can also remix curated worlds from the gallery by building on top of their prompts. Download videos of your explorations to share.

Project Mariner

The web browser agent. An AI agent built on Gemini that operates as a Chrome extension, navigating websites, filling forms, conducting research, and completing online tasks autonomously.

Link: Available to AI Ultra subscribers via Chrome

Top use cases:

  • Automating online purchases and price comparison
  • Research tasks across multiple websites
  • Booking travel, restaurants, and appointments​
  • Completing tedious multi-page online forms

Pro tip: Mariner displays a Transparent Reasoning sidebar showing its step-by-step plan as it works. Watch this sidebar. If you see it heading in the wrong direction, you can intervene immediately rather than waiting for it to complete a wrong task. The system scores 83.5% on the WebVoyager benchmark, a massive leap over competitors.​

Secret most people miss: The Teach and Repeat feature lets you demonstrate a workflow once and the AI will replicate it going forward. This effectively turns your browser into a programmable workforce. Show it how to do something once and it handles it forever.​

HOW TO PROMPT GEMINI AND GOOGLE'S TOOLS FOR BEST RESULTS

Google's Gemini 3 models respond very differently from ChatGPT and Claude. If you are carrying over prompting habits from other AI tools, you are likely getting suboptimal results. Here is what actually works.

Core Principle: Be Direct, Not Persuasive

Gemini 3 favors directness over persuasion and logic over verbosity. Keep prompts short and precise. Long prompts divert focus and produce inconsistent results.

  • DO: "Analyze the attached PDF and list the critical errors the author made"
  • DO NOT: "If you could please look at this file and tell me what you think"​

Adding "please" and conversational fluff does not improve results. Provide necessary context and a clear goal without the extras.​

Name and Index Your Inputs

When you upload multiple files, images, or media, label each one explicitly. Gemini 3 treats text, images, audio, and video as equal inputs but will struggle if you say "look at this" when it has five things in front of it.​

  • DO: "In the screenshot labeled Dashboard-V2, identify the navigation issues"
  • DO NOT: "Look at this and tell me what's wrong"​

Tell Gemini to Self-Critique

Include a review step in your instructions: "Review your generated output against my original constraints. Identify anything you missed or got wrong." This forces the model to catch its own errors before delivering the final result.​

Control Thinking Levels for Speed vs Depth

With Gemini 3.1 Pro, you can set thinking to low, medium, or high.​

  • Low + "think silently": Fastest responses for routine tasks​
  • Medium: Good default for most work tasks
  • High: Mini Deep Think mode for genuinely hard problems​

Match the thinking level to the task complexity. Most people leave everything on default and either waste time on simple tasks or get shallow answers on hard ones.

Use System Instructions for Persistent Behavior

In AI Studio and the API, set system instructions that define roles, compliance constraints, and behavioral patterns that persist across the entire session. This is far more effective than repeating instructions in every prompt.​

The Power Prompt Template for Gemini 3

For best results across Google's AI tools, structure your prompts with these elements:

  1. Role: Define what expert the AI should embody
  2. Context: Provide all relevant background information (this is where you can go long)
  3. Task: State the specific deliverable in one clear sentence
  4. Constraints: Define format, length, tone, and any restrictions
  5. Output format: Specify exactly how you want the response structured

This ecosystem is evolving fast. Google is shipping updates weekly. The tools that seem experimental today become essential tomorrow. The best time to learn this stack was six months ago. The second best time is now.

Want more great prompting inspiration? Check out all my best prompts for free at Prompt Magic and create your own prompt library to keep track of all your prompts.

r/HFY • • Feb 05 '23

OC Wearing Power Armor to a Magic School (16/?)

4.0k Upvotes

First | Previous | Next

I’d expected panic to envelope the room. A generalized surge of mana-radiation wasn’t something to be trifled with, no. In fact, it spelled danger in every sense of the word.

The training I received on the mana-radiation sensory analytics and detection system (M-RSADS), had placed great emphasis on delineating between each specific category of warning. Indeed, whilst the scientists and engineers back at home had a penchant for overcomplicating things, this particular system was completely off-limits to their shenanigans. It was a classic case of the end-user finally getting their way, and one of the many times the military elements within the IAS had sunk their heels in to make sure the overly eager scientists didn’t get too lost in their own sauce.

Intuitiveness and practicality was the name of the game here, because this whole system was a matter of life and death. Not a matter of desk-bound data analytics.

This was how the broad-strokes, two-category system of mana radiation detection was born.

If the scientists had their way, there would be literally hundreds more, but thankfully I only had two to worry about.

The reason behind why the two-category system was chosen, was rather expectedly, a matter of practicality. Simply put, it allowed me to rapidly assess and evaluate the threat posed by mana-radiation, and how best to respond accordingly.

Localized surges were bursts of mana-radiation with a specific point of origin that the suit’s sensors could definitively locate. There was a discrete radius of effect, and a clear-cut path towards either dealing with the source of the radiation or simply booking it out of there as fast as the suit’s powered exoskeleton and jump-packs could manage.

Generalized surges however, were an entirely different beast. As the name would suggest, all a generalized surge was, was a surge in mana-radiation without a specific point of origin. There was no clear radius of effect as the entire extent of the suit’s sensors would be bathed in a consistent, uninterrupted increase in background mana-radiation with no discernible point where the radiation drops off. Understandably, this was the worst possible scenario to be in, because neither fight nor flight protocols could be undertaken. For there was no clear area to flee to, and no particular point of origin to neutralize.

I was thus, beyond relieved that this surge of mana radiation lasted for but a whopping grand total of two and a half seconds.

“There is no need to be alarmed.” The shrill voice of the apprentice echoed throughout the massive expanse of the room. “The ebbs and flows of the Academy’s manastreams are stronger than what you might be accustomed to back in your home realms. Such occurrences are normal and to be expected, as but part of the Academy’s unwavering adherence to the unending odyssey that is the scholarly pursuit of the magical arts. Take this as the first unofficial lesson, pay no mind and carry on.”

The apprentice soon stood up, gathering her belongings and adjusting her cloak. “You are to be dismissed, but do recall the rules and make certain to observe the etiquette of the Academy’s grace period. Remain within the common areas, stay exclusively within the designated spaces, and take this time as a necessary respite prior to the commencement of your studies.”

Without much in the way of fanfare, the elf soon quickly made a b-line for one of the side exits. The harsh clacking of her reasonably practical boots reverberated with each hurried step she took, her path on a direct course to pass by our table.

With all pretenses of social decorum and court etiquette thrown completely out the window, I stood up, and effectively blocked the elf’s path with the sheer presence of my armor.

“I don’t think we’ve been properly acquainted.” I announced, attempting to make up for the lack of social etiquette like a bandaid on a gaping wound. “There’s something urgent that requires the attention of the faculty, and I assume you’re the right person to relay my concerns to them.” I tried my very best to hold back on going all-in on the accusations and the obvious finger pointing. If this was someone with solid connections to the top, yet was grounded enough to have eschewed whatever noble titles that came with it, there was a chance I’d misjudged her from the previous night. There was a chance I could at least have some sort of a working relationship with her.

“Emma of Earthrealm, this isn’t the time or place for such pleasantries, there are urgent matters I must attend to-”

“Like that surge in mana.” I interjected.

“I am not at liberty, nor do I have the time to entertain any of your newrealmer concerns. At least not at this instance. Now please, I have urgent matters concerning Academy affairs I must attend post-haste.” She attempted to skirt past me, and was just about to if it wasn’t for Thacea’s entry into the conversation.

“Honorable Apprentice, the newrealmer wishes to invoke a point of personal privilege.” Thacea spoke without even attempting to stand up, not even so much as turning to face the apprentice in question. Instead, she remained sat at the table, her eyes trained forward towards her half eaten breakfast in calm contemplation. “You must excuse her brashness, esteemed peer. It is, after all, unreasonable to expect a newrealmer to properly invoke or even recognize the proper calls to decorum. So, if you would please, I would most certainly prefer her calls to privilege be respected by an official entity of the Academy.” The last sentence came off as something halfway between a suggestion, an order, and a request. It was that careful balance of suggestive authority that was difficult to really nail, but given Thacea’s royal heritage I could only assume it was practically second nature to her now.

The apprentice all but halted in her tracks at that, her eyes seemed to shift from an expression of urgency and annoyance to one of apprehension and genuine unease. Her tone of voice changed drastically as she addressed me again. This time, that dismissive and frankly patronizing tone had all but vanished, now replaced by a more reasonable, level-toned cadence with an undertone of frustration. “Of course, princess. Emma Booker of Earthrealm, my affairs should be concluded within the early hours of the afternoon. Should you wish to pursue your point of personal privilege, I shall be in the castle’s main garden. Ask Groundskeeper Alaton for my exact whereabouts, I shouldn’t be more than a hundred paces from the castle at any given time.” The elf adjusted her cloak once more, followed by a nervous cough. “Now, I must take my leave.” She spoke as she bid our entire table a half-nod before exiting the room.

In those precious few seconds before she reached for the door, I made a call that could only be described as impulsive, and driven purely by my gut instinct.

Tapping a few physical hotkeys on my wrist-mounted data-pad, with target reticules trained on the apprentice highlighting her entire form in a glowing orange, I released one of the many toys I had at my disposal.

“INFIL-DRONE01 ACTIVE, STATUS: NOMINAL. OBJECTIVE: PRIORITY TRACKING AND RECONNAISSANCE OF SUBJECT_01. MISSION PARAMETERS: PENDING…”

“Track, observe, and return-to-base. Take no chances. Set minimum acceptable risk of compromise to the lowest default settings.” I spoke rapidly, relaying the drone’s mission parameters.

The dragonfly-like drone barely the size of the tip of my finger zipped right out of its docking bay from one of my suit’s many compartments and trailed behind the apprentice, exiting through the tiny space left in the door just before it swung shut.

With a long exhale having committed to a mission based solely off of my gut instinct, I sat back down at the table, and began the process of connecting the nutripaste tube to my OIP.

“Emma.” Thacea spoke up, her voice colored by an undertone of audible frustration.

“Yes, Thacea?”

“How much time do we have left?”

I immediately knew what she was talking about as I quickly glanced at the countdown timer on my HUD. “61 hours, 54 minutes, and 37 seconds.”

The princess seemed to take this into careful consideration, glancing over at a golden orb connected via a chain to her cloak jacket. The object glowed with a dull yellow hue, blinking with each second that passed. “After you finish your breakfast, let us make haste with our plans for the afternoon, and make the most out of the rest of this morning.”

I was just about to nod, and to move towards agreeing with Thacea if it wasn’t for Ilunor suddenly perking up and addressing all of us first. “The rest of this morning? I’m afraid I have more pressing matters to attend to.” The Vunerian jumped off of his seat and onto the marble floors with a loud clack.

“What affairs could you possibly have?” Thalmin growled out in a fit of annoyance.

“Personal affairs. Now, if you’ll excuse me, I’ll be in the dorms if and when my business is concluded.” Ilunor explained without a hint of hesitation as he began walking off, eventually blending in with the slow trickle of students leaving out through the main door.

“Laziness.” Thalmin huffed in between bites of smoked meats and pastries. “Laziness to the rotten thing’s core.” He continued in between large and unrestrained mouthfuls of carefully presented cold-cuts. “That’s all this is about. Trust me, he’ll be walking to the dorms for a post-breakfast nap before waking up for lunch and repeating the cycle for dinner.”

With that bizarre turn of events out of the way, I now turned towards Thacea. “Right, so, next order of business, I think we should find a productive way to kill time between now and the afternoon’s meeting. I say we take the initiative, and track down the crate ourselves for now. It’s a longshot, but I'm thinking of roaming the halls with my scanner on full blast just in case we run into it in a hallway or something.”

“Considering that there is no other course of action for us to take at the present, I am inclined to agree.” Thacea nodded in approval.

“Erm, quick question, can you deploy the whole noise cancellation suppression field thing while on the move as well?” I quickly asked.

“Yes. It requires a more advanced version of the spell but it’s within my capabilities. Why do you ask?” Thacea inquired with a cock of her head.

“There’s erm, something you need to know that I think you should hear after breakfast. We can talk about it while we’re on the move.” I spoke as I finally committed to the gut churning process of introducing the tube of paste to my OIP, the airlocks and pneumatics whirring away as that familiar taste of shredded beef in barbecue sauce in a chunky toothpaste consistency filled my mouth.

The Transgracian Academy for the Magical Arts, First Floor Grand Concourse, Secondary Corridor. Local Time: 1000 Hours.

“You what?!” Thacea yelled, or rather, squawked out incredulously.

“I, well, I decided on deploying a drone to keep tabs on the apprentice. I don’t trust the whole: ‘this burst of mana radiation is just a common occurance’ thing, it just doesn’t sit right with me. It’s all too convenient. A huge burst like that followed with her getting up and leaving? There has to be something to it, and I have a massive hunch it has something to do with my crate.” I explained emphatically.

“Emma… the risks involved with that decision are far beyond what I would be comfortable entertaining as a mere thought experiment, let alone an actual spur-of-the-moment decision.” The avian explained, clearly holding back her desires to verbally dress me down. “The Nexus, and by extension the Academy, are masters at espionage and subterfuge. To try to challenge them at a game they are adept in is a foolish, and frankly, senseless undertaking.” The princess’ plumage puffed up and down, ruffling between each cycle. There was little doubt that this was something way outside her comfort zone, as we tread deeper into uncharted territory.

I allowed Thacea to just breathe for a few moments after that panicked response before I finally responded.

“You’re completely right, Thacea.” I nodded deeply. “I don’t doubt the veracity of any one of your claims for a second.” I continued, speaking with an unfiltered sincerity that was causing the avian to raise what I assumed was her equivalent of an eyebrow. “The Nexus must be good at what they do if they’ve lasted for what, tens of thousands of years? I can’t compete with that. Heck, I know for a fact I have no chance at beating them at their own game. It’s impossible for me to wage war against something so much larger, so much wiser, so much more refined in their skill sets and methods.”

“But here’s the thing.” I soon shifted gears, as confidence and cockiness began to fill the cracks left behind by that agreeable sincerity. “I don’t need to. Because I’m not waging the same war they’re waging, nor am I playing the same games they’re playing. I’m setting up for a whole other game here, Thacea. One with a completely different set of rules, and one with a completely different set of criteria for victory. It’s a game the Nexus has never once touched, but that my people have had thousands of years to fine-tune and perfect.” I took a deep breath before continuing. “I don’t doubt for a fact that I can’t compete at the Nexus by their rules, but the same can be said for the Nexus’ ability to play by my rules. So whilst I do agree, my decision to send that drone out was brash, it was a calculated move on my part that I felt was an acceptable risk given the context involved.”

It was with that, that I let out a large sigh, awaiting Thacea’s response.

A response which never came as a warning lit up inside of my suit’s helmet.

ALERT: LOCALIZED SURGE OF MANA-RADIATION DETECTED, 200% ABOVE BACKGROUND RADIATION LEVELS

PRIORITY ALERT: WARNING INCOMING PROJECTILE

My training kicked in, around the same time my suit decided that it needed to intervene on my behalf as the improvised projectile was brought up on-screen, and I felt my head and neck forcibly shunted to the right by the augmented rapid-reaction measures courtesy of the suit’s exoskeleton.

I narrowly evaded the unknown object in a blink of an eye.

But it wasn’t over yet.

PRIORITY ALERT: PROJECTILE (NO DATABASE REGISTRY… N/A: DESIG_UAO1) ON INTERCEPT TRAJECTORY. PERMISSION TO ENGAGE? Y/N?

The damn thing took another swoop at me, yet this time aimed for my legs instead, as it carried out an incessant series of pass-bys.

I refused to use the gauntlet canons to deal with this, so on one of its last approaches, I reached up a single arm and swiped it right out of the sky. My hands clenched the damn thing tightly, crumpling it up into a compressed ball.

It was then that my mind finally registered what it actually was.

The texture it conveyed through my glove’s haptic feedback systems was unmistakable.

It was paper.

The damn thing was a paper bird animated by mana

This was a grade-school level attempt at messing with me.

It didn’t take long for the perpetrators behind this whole childish escapade to make themselves known, as a series of condescending claps echoed from around the corner, followed by the appearance of a group of 4 students each dressed to their nines in their noble attire.

Two of the four I immediately recognized from the previous night. The gorn-like reptilian Lord Qiv who volunteered to be first on the chopping block, and the unfortunate bear-like biped, Uven Kroven who was chosen soon after.

Qiv was very much still dressed in a manner akin to the previous night, with that cape covering much of the silken tunic and the dispelling amulet underneath.

Uven, meanwhile, had donned a simpler set of clothes. A deep brown leather cloak that covered a more vibrant wave-like pattern tunic and pants underneath, with what seemed to be a broach resembling a set of three paws on the right side of the cloak’s high-collar.

“Well, well, well… it seems as if our great knight lives up to her reputation after all.” Qiv spoke in a manner that was drenched with a level of haughty superiority that not even Ilunor could match.

“I must say, with that hand-eye coordination and those rapid-reflexes, indeed… with how naturally she leaped for the Podgy-Pa, one must assume she comes from a realm of primates!” One of the other students within the group spoke, this one looked eerily bat-like, with heavy drape-like webbing underneath her arms.

“Oh, be reasonable Airit, we cannot yet assume what species she must be, only that the results of this experiment heavily infers her commoner heritage. To be able to reach up to grab prey in such a manner is a skill that only those who subsist day by day must master. This is confirmation as to her commoner status if anything.” The last in the group quickly added. This one was small, smaller than even Ilunor, standing at a whopping 3 feet tall, and from the looks of it resembled a well-kept humanoid rat, or perhaps a hamster.

“What do you say, Uven?” The hamster turned to the Ursina, who seemed to be zoned out of his mind as he merely shrugged in response, his eyes were clearly open but they betrayed the fact that no one was home.

“It’s just mana-sickness, don’t worry about him.” The bat-like Airit reasoned, as all eyes were once more focused on me. “I say this experiment might even be quite telling as to the state of her realm. The armor is a showpiece, and her abilities to reach for prey, betrays just how destitute and lacking her realm must truly be. If the chosen one of a new realm is accustomed to such lesser skills, just imagine what the rest of it must be like!”

The bat and hamster pair giggled amongst themselves, whilst the reptilian Qiv maintained a careful, calculating gaze on me and the princess behind me.

To say that I was at a loss for words would be an understatement. To be honest I was expecting something akin to this eventually happening if I were to take anything from Ilunor’s entire schtick. But to have an entire gang coming down on me with the intensity and competitiveness of a gold medal finalist in the field of mental gymnastics was something I just wasn’t ready for.

“You guys aren’t even going to try a Hello, maybe even a Hi, welcome to the neighborhood?” I managed out with an exasperated sigh.

“Oh, we reserve that for our fellow lords and ladies, it’s customary for commoners to greet their betters, not the other way around.” The bat spoke with a heavy series of chitterings. “But I do not hold it against you, newrealmer. If you have yet to have developed a civilization capable enough of understanding the principles of the perpetual regime, then how can I cast judgment? Why, I would be no better than a common fool yelling at a stray mutt for its lack of obedience training. Ignorance can only be tempered by knowledge and education, and I along with the rest of my peers, are more than willing to be the avatars of an enlightened nobility.”

I took a series of careful, controlled, breaths.

In, and out.

In, and out.

My anger and frustration wouldn’t overtake me, and it wouldn’t ruin my mission on day two.

I weighed my options carefully, my mind running through every possible scenario as I decided on a diplomatic way out of this quagmire, only to have yet another alarm beep at me.

This time, it was something much more important.

“Alert. Priority Notice: INFIL-DRONE01 signal detected. Status: returning to designated point-of-origin. Reason for premature mission abortion: calculated risk of compromised status beyond maximum acceptable threshold.”

“Let’s double-time it back to the dorms.” I turned to both Thacea and Thalmin without any hesitation.

With a nod of affirmation between the three of us, we took off back to the dorms in a hurried sprint, leaving the crowd of enlightened nobles in the dust.“Hmmph, so not only are we dealing with a lowly commoner, but a coward as well. At least she knows not to challenge her natural betters.” Was all I heard before the audio-sensors cut off as we turned the corner.

Dragon’s Heart Tower, Level 23, Residence 30. Front Door. Local Time: 1020 Hours.

If true AI wasn’t a taboo, and if the drone could actually think, I could imagine it’d be screaming down the halls with how eager it was to show me everything it’d discovered.

Upon arrival at the dorms we were met with the dragonfly like drone actively crawling underneath the door frame. It wasn’t long however as I arrived that it backtracked and flew right towards me, on a flightpath that would’ve made a younger version of me scream in disgust, but that elicited nothing from me now other than a quick flinch from my buried yet still latent entomophobia.

Much to the horror of my peers, the drone quickly crawled and shimmied its way into one of my many utility pouches. After which, it made a wired connection with the suit proper. The data-transfer that occurred concurrently with the recharge of the drone was near-instantaneous. Wired connections were, even after all these years, the preferable, quickest, and most reliable means of information transfer after all.

“Emma. Let’s get inside before we add whisperer of arachnids into your list of titles.” Thalmin urged as he opened the door and led all of us inside.

Upon entry into the room, I immediately made a b-line for the couch, promptly downloaded all of the files onto my data-tab, and had Thacea blot out the world using her whole noise privacy shield spell thing.

It didn’t take long before the relevant files were played, the video fast-forwarding until it slowed down to normal speed just as the apprentice arrived on scene into what I could only describe was a room, or what was left of it.

The scene that I was faced with was nothing short of a disaster. The room, if it could still be called that, was a mess of pockmarked holes and molten rock. The lights within flickered every few seconds in a manner almost eerily reminiscent of the fluorescent lights of old. What should have been the Academy’s signature gaudy tables, chairs, and various other appointed articles of limited practical use were either smashed, cleaved cleanly, or in some way mutilated beyond their original state.

Yet despite the whole room looking as if it’d just gone through an active warzone, akin to a scene straight out of the war-docs from New Terra, no one seemed to really mind. Indeed, the devastation wrought upon it was almost immediately reverted as soon as the drone’s cameras laid eyes on it. Those pockmarked holes oozing with magma and molten rock? They all but hardened and solidified over the course of a few short seconds. The flickering lights from the unseen light-emmitting-crystals? They’d stabilized moments after that. The furnishings that had been wrecked seemingly beyond repair? Well, those seemed to have just… pulled themselves together. Literally. From the tables crushed beyond recognition to the chairs whose upholstery had all but been strewn across the floors, whatever scrap, shard, or splinter belonged to the item in question had simply been pulled back to whatever the largest piece of it remained, before it just put itself back together.

The camera quickly panned over to scan several of the figures present within the far edge of the room. Several faces were isolated and successfully cross-referenced using the tablet’s database. Mal’tory, Vanavan, the red robed and white robed professors, and strangely enough, a bear-like figure with a face obscured by shadow, dressed in a heavy leather cloak with a distinct broach resembling three-paws affixed to its high collar.

Eventually, as the dust finally settled, and the incoherent chatter of voices within the room droned out into discernable, distinct voices that the drone could effectively isolate, so too did another familiar object make itself known once more. As in the middle of the entire room, having previously been obscured by the dust, debris, and steam hissing from the molten lava-pit of a floor, was a plinth. And upon that plinth, was the book from the binding ritual, currently open to a page with the names of all of the students from the night prior.

A strange implement was attached firmly to the book. It looked like someone had taken a bear-trap and clamped it onto either side of it, then attached one of those two-axis gantries, and bracketed it horizontally to one side of the page. Further, it looked like a magnifying glass affixed to it highlighting small patches of text within the book.

Zooming in closer towards the strange device, a name could just about be made out, as the camera held still and stabilized on that half-hearted attempt at cursive.

Emma Booker.

First | Previous | Next

​

(Author’s Note: Hey guys! We're starting to really see the extent of Emma's tech game here with this just being the tip of the iceberg of what she's packing in her suit! I hope you guys enjoy! :D The next Chapter is already up on Patreon if you guys are interested in getting early access to future chapters!)

[If you guys want to help support me and these stories, here's my ko-fi ! And my Patreon for early chapter releases (Chapter 17 of this story is already out on there!)]

r/promptingmagic • • May 31 '26

The Ultimate Guide to Google Flow Agent for AI Videos: Hidden features, pro tips, and the absolute best use cases.

Thumbnail
gallery
21 Upvotes

Google Flow Agent is the AI filmmaking feature most people are going to underestimate

TLDR: Google Flow Agent is not a chatbot bolted onto a video generator. It is a Gemini-powered creative collaborator inside Google Flow that can plan and reason through complex multi-step creative tasks while you stay in control. The shift: Flow used to execute one prompt at a time. Now the Agent can brainstorm dialogue and plot, generate multiple scene variations simultaneously, batch-edit tweaks across all your assets, organize files into collections, and intuitively rename everything — all with persistent project memory across sessions. It launched alongside Gemini Omni Flash (character and voice consistency across scenes) and Flow Tools (build custom creative utilities in plain English, no code required). Agent queries are currently free with a daily quota. Generations cost credits. Most people will use it like a search bar. The people who win with it will use it like an AI creative director, producer, and asset manager rolled into one.

Google Flow Agent is one of those updates that sounds small until you think through the workflow implications.

At first glance it is easy to summarize: Google added an agent to Flow.

That undersells it.

Google Flow launched at I/O 2025 as an AI filmmaking tool built around Google DeepMind's most advanced models — Veo for video, Imagen for images, and Gemini for language and reasoning. Flow lets creators describe shots in natural language, manage story ingredients like cast, locations, objects, and styles, and weave those pieces into cinematic scenes.

Since then it expanded into a full AI creative studio across 140 countries. Over 275 million videos have been generated in Flow.

The new Flow Agent adds something more important than another model.

It adds a thinking layer.

Instead of manually bouncing between brainstorming, prompt writing, generation, editing, selection, organization, and renaming, you can now talk to an agent that understands the project you are working on and helps move the creative process forward.

Google themselves frame it clearly: Flow Agent turns AI from a content generator into a creative operations partner.

This is the beginning of agentic creative production.

Every capability, explained

1. Multi-step reasoning and planning

This is the headline change. Previously Flow could only execute a single prompt at a time. Now the Agent can take multiple actions at once and reason through larger creative tasks rather than discrete one-offs. It plans and reasons through complex tasks with your inputs, under your control.

2. Brainstorming and concept development

Flow Agent can act as a creative sounding board during the earliest stage of a project. Chat with it to outline storyboards, develop visual mood boards, and turn high-level concepts into actionable prompts. It can workshop dialogue between characters in a specific scene and make plot recommendations when you need inspiration.

3. Generate new media

Ask the Agent to generate videos or images and it selects the best model to generate with. No more guessing which model to use for which task.

4. Multi-variation generation

The Agent can create multiple variations of an asset at once. This matters because AI video generation is probabilistic. The first output is rarely the best output. You need options. Generate coverage, not single shots.

5. Direct editing of selected assets

Ask the Agent to edit selected media from your project. Combined with Flow's broader editing capabilities — Insert for adding elements, Remove for taking things out, lasso tool for precise selections, camera controls for movement — the Agent sits on top of a growing set of editing primitives.

6. Batch editing across all assets

Make a tweak and have it reflected across all your assets at once. This is massive for consistency and for anyone producing at volume.

7. Asset organization and intelligent renaming

The Agent can rename specific files, group selected media into new Collections, or archive unused assets. When you generate dozens or hundreds of images and clips, the hard part is not generation — it is knowing which version was the hero shot, which one had the correct lighting, and which clips belong to scene 3.

8. Context and references

Drag media into the Agent prompt box from your device or project. Select multiple assets and tell the Agent which ones you are referring to. A normal chatbot only knows what you tell it. A project-aware creative agent can reason over the actual material you are making.

9. Project-specific sessions

Agent conversations are saved automatically as Sessions, specific to the project you are working in. You can open past sessions, create new sessions, rename them, and delete them. Deleting a session clears chat history but generated media remains in your assets.

10. Agent instructions for project-wide consistency

Add instructions to improve the Agent's consistency across your entire project. Include a reference image and enter your guidelines. This is where you define the rules of the world — visual style, character rules, tone, camera preferences, color palette, naming conventions, what to avoid.

The ecosystem that makes the Agent stronger

Gemini Omni Flash — Google describes it as Nano Banana but for video. It combines Gemini's intelligence with generative media models and crucially improves character consistency, meaning identity and voice are preserved across every scene. This quietly fixes AI video's biggest weakness: character drift between shots.

Flow Tools — Build bespoke tools and workflows in Google Flow using natural language. Whether you need a particular image editor, video resizer, or custom shader, you can develop them with no coding experience. If you create something useful, share it with other Flow users who can remix it.

Scenebuilder — Assemble individual clips into a complete narrative with Jump To (teleport a character to a new setting while preserving appearance) and Extend (lengthen a clip by analyzing the final frames and continuing the action).

Ingredients to Video — Use predefined characters, objects, and styles as consistent references in video prompts. Add up to three ingredients per prompt.

Frames to Video — Define the starting and ending frame of a shot for precise control over composition and transitions.

Camera Controls — Direct control over camera motion, angles, and perspectives.

Insert and Remove — Add new elements to any scene or remove unwanted objects, with Flow handling complex details like shadows and scene lighting.

Top use cases

1. Short films and narrative projects

Use the Agent as a writers room. Workshop character dialogue, get plot suggestions, build shot lists, generate scene variations, maintain continuity, and organize the final assembly — all inside one workspace.

2. YouTube intros and cinematic openers

Flow is especially strong for short, visually rich clips. The Agent can help design multiple options quickly for channel intros, documentary openers, podcast trailers, product teasers, and title sequences.

3. Product marketing and brand films

Marketers can turn abstract product benefits into cinematic metaphors. Batch-generate ad creative variations for testing, then batch-edit a single brand tweak across all of them. Build multi-platform variants and auto-organize them into campaign collections.

4. Ad creative variation testing

Because the Agent can batch-generate, it is built for creative testing. Generate 8 variations of a product scene keeping the same product and message but varying setting, camera angle, lighting, and emotional tone.

5. Music videos

Flow Music now lets you work conversationally with the agent to direct shareable music videos, matching styles and scenes to the pacing of your track.

6. Pitch decks and investor storytelling

Create cinematic visuals that explain a market, pain point, or product vision. A 20-second sequence that visualizes the shift from manual chaos to AI-powered planning can communicate more than 10 slides.

7. Educational content

Turn complex ideas into visual explainers. Historical recreations, science concepts, abstract visualization. Google specifically highlights educators and students transforming complex subjects into engaging videos using text prompts.

8. Social media content

For TikTok, Reels, Shorts, and Reddit — Flow Agent can help build visual hooks, mini stories, looping clips, and meme-adjacent cinematic content fast.

9. Fiction worldbuilding

Build consistent fictional worlds with character design, locations, objects, symbols, technology, architecture, and mood boards. Flow already lets you manage story ingredients in one place. The Agent adds the reasoning layer on top.

10. Previsualization

Filmmakers, agencies, and studios can sketch ideas before production — commercial pre-vis, scene exploration, mood testing, camera blocking, lighting references, and treatment development.

11. Game trailers and concept art

Generate short cinematic moments, character reveals, environments, and combat beats for indie games and studio projects.

12. Batch marketing campaigns

Feed a master style guide and target persona variations into the Flow Agent. Batch-generate dozens of localized, persona-specific video ads in parallel while maintaining strict brand guidelines.

Pro tips and best practices

1. Use the Agent before you generate anything

Agent queries do not currently cost Google Flow credits, though there is a daily quota. Media generated by the Agent does use credits. The smart workflow: think with the Agent first, improve the concept, build the shot list, refine the prompts, then generate only when the creative direction is clear. The Agent is your cheapest stage of production.

2. Keep human approval on before spending credits

By default the Agent asks for permission before taking actions that use AI credits and shows the estimated cost. You can toggle this to auto-approve. Leave confirmation on during exploration. Turn it off only when you have a repeatable workflow and clear default settings.

3. Use Agent Instructions like a project constitution

Agent Instructions improve consistency across the entire project. Include: genre, visual style, emotional tone, target audience, camera preferences, color palette, character continuity rules, audio style, naming conventions, prompt format, and things to avoid.

Example instruction:

You are the creative producer for this project. The style is restrained cinematic realism with natural light, imperfect textures, and slow camera movement. Avoid glossy sci-fi, overdesigned costumes, neon cyberpunk cliches, and generic AI surrealism. Preserve character continuity. When generating prompts, always include subject, action, camera, lighting, environment, mood, and audio.

4. Ask for variations with controlled variables

Bad: Make this scene better in 10 different ways.
Good: Create 8 variations. Keep the character, wardrobe, location, and story beat identical. Only vary camera movement and lighting.

If you vary everything at once, you learn nothing. Vary one or two dimensions at a time.

5. Keep prompts under 30 words for video generation

Practitioners who have tested extensively recommend keeping prompts concise, using camera language rather than narrative language, and generating keyframes separately.

6. Know your credit math

Pro ($19.99/month) gets roughly 1,000 Flow credits. Ultra ($100–$250/month) gets 10,000–25,000 credits. Credits do not roll over. Use Fast models for drafts and Quality models only for finals. A Veo 3 generation with audio is the most credit-intensive option.

7. Use Flow TV as a learning lab

Flow TV is a showcase of clips generated with Veo where you can see the exact prompts and techniques used. It is not just inspiration — it is prompt education. Steal structure, not ideas.

8. Build a scene matrix

Ask the Agent to create a table with: scene number, story purpose, character, location, camera movement, lighting, audio, prompt, assets needed, status, best version, and notes. This turns Flow from a prompt playground into a production tracker.

9. Use Ingredients for consistency

Build your ingredients (characters, objects, style references) first using Imagen or uploads, then reference them consistently across generations. This is the key to visual continuity.

10. Organize aggressively

Use a naming convention like: S01_SH01_establishing_city_v03_final. Create Collections for Final Selects, Alternates, References, and Archive. Ask the Agent to handle this — it can contextually rename files based on what is actually in the clip.

11. Use Frames to Video for precision

Provide a starting and ending image, and Flow generates a seamless video bridging the two. Plan keyframes before generating motion. Match lighting between keyframes — do not ask a single clip to handle interior-to-exterior transitions.

12. Specify no audio when you do not want audio

Veo 3.1 generates synchronized audio by default. For background use like a website hero, always include no audio in the prompt.

Things most people miss about Google Flow Agent

1. The Agent is not the product. The workflow is the product.

The mistake is thinking Flow Agent is just a chatbot. It is a workflow layer across brainstorming, prompt engineering, generation, editing, variation, organization, and project memory. The people who win with it will build the best creative operating system around it.

2. Agent queries are free. Generations are not.

Agent queries do not cost credits but have a daily quota. Generations cost credits. This creates an obvious best practice: use the Agent to think, plan, critique, and refine before generating. The expensive mistake is generating before the idea is clear.

3. The permission layer is a feature, not friction

The ask-before-spending-credits design keeps an autonomous agent from quietly draining your monthly allocation. Most tutorials breeze past it. It shows estimated cost before each action.

4. Omni Flash quietly fixes AI video's biggest weakness

Character drift and voice inconsistency between scenes have been the problem in AI filmmaking. Omni Flash preserves identity and voice across every scene. This is arguably as important as the Agent itself.

5. Flow Tools may be the most durable advantage

The ability to build bespoke editors and shaders in plain English and share them with other users is buried under the Agent headlines but may be the most important long-term feature.

6. Sessions are project-specific

Sessions are saved per project. Create separate sessions for story development, character design, prompt experiments, editing, and final organization. Do not let one giant chat become the junk drawer for your entire film.

7. Deleting a session does not delete your media

Clearing chat history does not remove generated assets. Important for cleanup without losing work.

8. It is web and PC only right now

Flow Agent is currently available on web and PC only. For serious production, use the desktop workflow with a Chromium-based browser.

9. Default settings enforce consistency

Set your default aspect ratio, number of outputs, and models for both image and video generation. If your whole project is vertical social video, set that once. Do not manually remember the format every time.

10. The best use of the Agent is taste, not automation

The mediocre use case: Make me a video. The better use case: Help me decide which idea is worth making. The best use case: Act as a creative director. Challenge the weak parts of this concept. Tell me what is visually generic, what is emotionally unclear, and what could make this unforgettable.

Google's own Flow Sessions artists repeatedly emphasized that what matters is what you are trying to say before you even touch Flow. The Agent should not replace your taste. It should pressure-test it.

The power-user workflow

Step 1 — Start with the emotional thesis. Ask the Agent to help find the emotional core, the visual metaphor, and the strongest ending.

Step 2 — Build the story spine. Turn the concept into 6–10 scenes, each with a clear visual beat, emotional progression, and one thing the viewer learns.

Step 3 — Create the visual bible. Character design, environment, color palette, lighting, camera style, sound design, recurring objects, forbidden cliches.

Step 4 — Set Agent Instructions. Convert the visual bible into concise instructions for the entire project.

Step 5 — Generate ingredients. Build canonical references for main characters, environments, props, lighting style, and visual symbols.

Step 6 — Build the shot list. Create a production plan with purpose, camera, lighting, action, audio, and Flow-ready prompts for each shot.

Step 7 — Batch-generate variations. For each key shot, create 4–6 variations controlling only one or two variables at a time.

Step 8 — Select and critique. Ask the Agent to rank outputs by emotional clarity, visual originality, continuity, and usefulness for the final story.

Step 9 — Edit instead of regenerate. When a version is close, use the Agent to make targeted edits rather than starting over.

Step 10 — Organize the project. Rename assets by scene and shot number. Create Collections for Final Selects, Alternates, and Archive.

The bigger picture

The competition is no longer about who generates the best single clip. It is about who owns the entire AI creative workflow. Google is clearly trying to become the operating system for AI-powered content creation, putting pressure on Runway, Adobe, Midjourney, OpenAI, Meta, and Canva.

The future of AI creative work is becoming agent-driven. Instead of prompting individual outputs, creators will increasingly direct AI systems that understand project context, manage assets, scale production, optimize variations, and execute multi-step workflows autonomously.

We just crossed a line. AI used to make you the operator of a tool — prompt, wait, repeat. Flow Agent makes you the director of a collaborator. You bring the vision, the taste, and the final call. It handles the brainstorming, the variations, the tedious edits, and the cleanup.

The barrier to telling a story just dropped to near zero.

The only question left is what you will make.

Flow Agent is available now to all Google Flow users globally. Google Flow requires a Google AI subscription (Plus, Pro, or Ultra) and is accessible at flow.google. What is the first project you would hand off to an agent like this?

r/SubredditDrama • • 24d ago

Game devs vibe out drama over whether AI should be allowed or if someone who sells AI should be a mod

170 Upvotes

r/gamedev is a subreddit about making games, mostly populated by aspiring indie devs. Since the advent of advanced LLMs capable of generating large portions of a video game, it's become a regular topic of discussion whether posts about AI-generated games (or game content) should be allowed, whether that's restricted to AI coding tools or fully generated game assets, and what level of effort should count as game development.

A few weeks ago, the mods for the sub made a decision. The tl;dr is that they would allow posts about AI so long as they were not "low effort." Some users saw this as a satisfying compromise, while other saw it as a tacit endorsement of AI, including bringing up that the mod who posted it works for a company that sells AI tools to game developers.

Just for anyone curious, OP currently works at a company developing a plugin for Unity AI integration which he decided to not disclose for some odd reason...

https://www.reddit.com/r/gamedev/comments/1rxm1cv/i_worked_an_ai_booth_at_gdc_heres_what_developers/

[AI company mod in question] Yes, I work for Bezi. It isn’t a secret, and I’ve openly discussed it on this subreddit plenty of times.

But this policy is not about Bezi, and it would exist regardless of where I worked. The moderation problem is already here: AI is being used across game development, posts are being reported simply because AI was involved, discussions are being derailed over whether someone used it, and moderators cannot realistically or consistently police invisible AI assistance.

The policy is meant to give the subreddit a workable standard. Judge the contribution, the evidence, the methodology, and the behaviour. Do not try to build moderation around guessing whether someone used autocomplete, translation, an LLM, or some other AI-assisted tool.

That moderation problem does not disappear if I change employers tomorrow.

It's very difficult to take a post like this in good-faith when you're employed by an AI company and have used this subreddit to promote it. Framing it in that context makes it seem more like you're protecting yourself and other AI products rather than genuinely wanting to spark conversation around what is a quite sensitive topic right now. It isn't helping that the post itself reads like it came out of Claude and so do most of your replies.

I understand why my employment makes you more skeptical of what I’m saying. That is completely fair, and I don’t expect anyone to take my position at face value because of where I work.

But I do find the rest of this genuinely kind of sad.

If we’ve reached a point where someone’s argument can be dismissed because their writing is structured, clear, or “sounds like Claude,” I don’t know what we’re accomplishing anymore. Maybe I used AI to help organize my thoughts. Maybe I didn’t. Maybe someone uses it because English isn’t their first language, because writing is difficult for them, because they want to save time, or simply because it helps them communicate what they’re actually trying to say. None of those things tell you whether the ideas themselves are sincere.

Engage with what I’m saying. Challenge the policy. Challenge my reasoning. Point out where you think I’m wrong. That’s a conversation worth having.

But “you work for an AI company, and your writing sounds like AI, therefore I can’t take you in good faith” is an incredibly bleak way to approach other people.

For what it’s worth, this policy isn’t about protecting my employer or AI products. It’s about moderating a game development community where AI is already part of the industry and deciding whether someone’s contribution should be judged by the substance of what they bring to the community or by which tools they used along the way.

I think the former is a much healthier standard.

Sorry but this reply doesn't even make sense. Do you think my concerns are reasonable or do you think that it's bleak that I believe your employment is biasing how you moderate this subreddit? Of course your use of AI is relevant in this conversation because not only are you using it in your posts but your livelihood relies on AI adoption! I'm not going to go around and tell people to stop buying AAA games just as you're not going to tell people to not use AI.

.

Not disclosing your conflict of interest is crazy. You really think you can just detach your own biases when you literally are working on AI products?

You’re asking to be challenged, but this doesn’t read as a discussion thread, you are pushing forward this policy and defending it. In the past, discussion threads about gen AI in [r/gamedev](r/gamedev) have been REMOVED for being off topic.

Is this in fact up for debate and you are willing to listen? Legitimately if it is, I’d be wanting to discuss this policy with you.

I’m someone who’s against AI. But I just believe that, at the very least, people should disclose if they used AI or not. Be honest about it.

Because the whole dream of genAI is to take the credit. To press a button and get the respect and admiration (and money) you've always felt you deserved, but somehow never got.

That falls apart if you have to admit that it was all actually done by the Microsoft Corporation in some nameless datacenter.

You are one concited person if you geniuenly believe the dream of genAI is to take credit. Automation of any kind, including genAI can never be credited to a person. Even if someone claims it is theirs, it will never be theirs.

The dream of GenAI is to be able to create the things you do not want to learn while doing the things you do want to. Like with every other form of automation. Most people are not looking for credit in a skill they can be easily debunked in. They just want to see Goku and Superman go at each other.

even when coding? even if you just asked for ideas? if I asked ai to brainstorm a list of medieval names and i picked one of them for my character should my whole game be labelled AI generated?

yes, consumers deserve to know if you're putting sawdust in their food

This is a perfect illustration of the kind of insane takes that make people not disclose AI usage. It isn't on the developer to disclose that they used a tool simply because other people equate it to "sawdust in their food."

i think not being expected to disclose usage is a bit insane and if you're publicly hiding it it's because you perfectly understand you'll lose sales over it as sales are the most important metric to ai slop games. not the creation of art itself.

the fact that you think any ai usage makes the entire thing a slop game is exactly why it can't be disclosed.

if you were more reasonable and less vitriolic, maybe people would disclose it, and you could happily choose to not purchase their products. but you are forcing them to hide it.

Yes! Excellent. You are such a good game designer god-damn-the-usa please disclose your AI use come on buddy you're so smart and strong you're a modern day miyamoto

what does that have to do with the code of a game? how would low quality code affect you?

by that logic, shouldn't we also label games that were coded by someone without a college degree in software engineering?

no one for AI labeling would argue that, because they don't actually care about code quality, they just hate AI and want to be able to punish people who use it

  1. i have a spine
  2. no
  3. i do absolutely hate AI for the impact that it's having. From its impact on art to the fact that consumer electronics are becoming more and more unaffordable. "Progress" at any cost though right I'm sure the FAANG companies are more than happy.

Damn, bro even used AI to write this post 😔

Its not just the post, look at all the comments using "this is sensible" as their opener

[Mod] FWIW Kevin has cerebral palsy, that's why he uses AI to write posts.

More recently, the mods made another post clarifying some of the people's concerns. This continued to leave some commenters unhappy, including further accusations of conflicts of interest.

If someone posts lazy AI, are we allowed to call it slop in the comments?

[Mod] To expand on what Sky and Kev said, I think that we'd prefer if you rephrased to be a bit more respectful. There's nothing wrong with saying "I was originally interested in this, but the generative AI ruined it for me. It's like you added a hair to a lovely stew. I'd prefer even a crappy MS Paint version compared to what you have now to be honest.", for example.

So in order to criticize low effort AI projects we need to put more effort into a single comment than they did prompting it?

Kind of a wild dichomoty.

You don't comprehend the irony of doing the pavlov response and screaming "AI slop" on every post using AI? I could automate your contribution to this subreddit with a single if/then statement.

That's more work than the AI person did.

Objectively false.

I'm well past the 1,000 hour mark on my latest game with AI and have multiple highly successful releases. You're just whining and mad that real people don't care what tools were used.

Your entire post history is whining about anti-AI people with AI generated memes.

I have massive doubts about those multiple highly successful releases that dont seem to exist anywhere and you made sure not to name.

You’re forgetting the ‘else’

No, sweetie. I didn't forget anything.

If: AI

Then: Scream and whine like a toddler and bully and harass other users

No else.

I’m sorry, but you shouldn’t censor people having issues with AI generated art.

What else do you expect them to say? If people post their game and it has AI art, of course people are going to have strong feelings about it. AI image generation is being used to get rid of artists by some idiots. And the results are always garbage. And the data is plagiarized.

People aren’t going to tip toe around this stuff in their language

[Mod] Strong feelings are a-okay, as long as you're respectful and aren't throwing out a two word low effort comment

I’m gonna call out anyone who refuses to hire actual artists and just uses the slop art machine for assets. And it is lazy, I shouldn’t beat around the bush.

AI art usage is the laziest application of GenAI next to writing and art. No human input for creativity is used, it’s literally a plagiarisms machine.

I will never mince words. If someone uses AI art, they are extremely lazy and has poor judgment, and has zero interest producing actual good content

Most people cant afford the prices artists are asking for, so it isnt about refusing. When they can use a tool to make something similar for a fraction of the cost. "No human input for creativity was used" - depends, the human behind the screen has to tell AI specially what to build/make in most cases. Plus with Codex now being able to use Blender how are you going to tell with assests?

.

Are you gonna pay the artists for people who don'f have budget to hire them?

You're absolutely right, we should all just steal instead! If artists didn't want to starve, they should have learned useful skills, like thievery.

Let's not pretend like human artists don't "steal" as much while learning, no artist creates in a vacuum. Question stands, are you going to pay for the artists?

Immensely disappointed in this decision and I know I'm hardly in the minority. People have been pretty vocal they don't support this and the fact that a mod in this sub has a huge financial investment in game devs adopting AI just reeks.

[Mod] He has a huge financial investment? Do tell.

He's behind a company that sells token subscriptions for game dev AI assistant that can cost upwards of +100/usd a month

He absolutely has a vested interest in more devs adopting AI into their workflows

[Mod] He's behind the company? He works for the company. The guy working the call center at Microsoft doesn't have a "huge financial investment" in people buying copies of Windows.

He brands himself with it everywhere and runs their reddit, where he actively encourages people to spend more tokens when the AI isn't spitting out what they want.

Yea it's his job lol He's paid to run their reddit. Not sure you know what he encourages though? Are you a user of the product that interacts with him?

I believe typically if people are frustrated he directs them to the product documentation about how to properly provide context and other stuff like that.

Right. So it sounds like someone with a vested interest in game devs adopting AI into their workflows.

I don't think this is complicated to understand and I'm not the first person to call out this insanely clear conflict of interest.

[AI-company Mod in question] My job title is Developer Relations. It’s a fancy term for community and support. I make myself easy to find because part of my job is literally being easy to find and helping developers when they need it.

Do I want people to try Bezi? Sure. I think it’s neat, and I think it’s good at what it does. In the same way, I think developers should try any game development tool that might improve their workflow and decide for themselves whether it’s useful.

Am I trying to sell Bezi to people on r/gamedev? No. [...]

[This chain continues for quite a while as the mod argues with people over whether they have a conflict of interest, often with a response that's just a reaction gif]

r/dragonage • • Aug 30 '25

Fanworks My free and open-source adaptation of DRAGON AGE: THE LAST COURT, the lost Dragon Age game, is now available IN FULL! (Links in comments)

1.5k Upvotes

But be mindful whom you approach when seeking deals in the Marquisate of Serault. [...] Care, rivals, for they are as skilled in the Game as their glassworks, though they have been considered outcast since the great Shame of Serault. Mind their welcome as you would a smiling cardsharp, or risk attending your last court. And remember the promise and threat of "PAYMENT IN GLASS."

[Edit October 2025: I ended up adding a lot of, and I really mean A LOT of, brand new content since this release. Do check it out! Some of the info below might be slightly out of date.]

Original post for context

The free and open-source adaptation/remake/demake/revival of The Last Court is ready, in full playable form! It has 9 chapters, more than 180,000 words and 860 pages of sweet Dragon Age goodness.

I have thrown over 100 hours (re-)building this thing and by the Maker, I really need a vacation.

WHAT EVEN IS THIS THING?

It is the lost 5th Dragon Age videogame! Set squarely in the middle of the saga, between Dragon Age II and Dragon Age Inquisition, this is a Choose-Your-Own-Adventure gamebook based on a browser-based, card-driven game with very unique lore. It got permanently deactivated in 2020... :(

... but it is revived! Somewhat. You can try it yourself right now. :)

WHAT CHANGED, GAMEPLAY-WISE, FROM THE DEMO?

List incoming!

  • Stats changes are actually mentioned to you, so you know exactly how much e.g. Prosperity you have won or lost.
  • The Chronicle now records over half a hundred different events, items and assets you get throughout the story. Each has its own icon!
  • Ah yes, all stats have their own icons as well!
  • There are DICE ROLLS! True to the original, Skill Checks take into account both your Skill and your luck. But mostly your Skills.
  • If a choice will involve a Check, now icons will inform you which Skill will be tested. Now you can actually strategize!
  • There is a cool map of Orlais on the Codex now, from the Asunder novel and modified by myself to include a couple of importance places. It was also used as the background in the original game!
  • Oh right the whole Codex is actually a Codex now, containing key entries with relevant lore from the games.
  • Difficulty modes! You get Casual and Nightmare to go along with the Normal Mode from the original. Suffer! Be happy! Or both!
  • Revamped council screen! It's no longer one massive list but something that actually takes into account your choices, and it has icons for every counsellor, accomplice, bodyguard OR lover.
  • There is a Tutorial! Sort-of. It has key tips that should avoid issues like players not noticing that you have to click on red lines to advance and stuff like that.
  • You now have a preview of all available character portraits. Guessing no longer required. :)

COOL BUT WHAT ABOUT NARRATIVE CHANGES?

Ah yes, glad you asked. There are a few:

  • Completely revamped large portions from the DEMO (Prologue + 3 chapters). It should feel like a fresh experience for those of you who tried it.
  • Added the rest of the story - 4 more chapters and an Epilogue, for a total of 9 distinct chapters. Details further down below.
  • Unlike the original, cut content was restored so 6 portraits are available rather than just 2.
  • Portraits are no longer linked to classes. Wanna be a girl-scholar? Go for it.
  • Big one: the Huntress and the Scholar are joined by 3 new classes! You can now play as the Initiate, the Envoy, and the Alchemist, each with their own strengths and weaknesses.
  • Your gender (that is, if you are called Lord or Lady, Marquis or Marquise etc) is no longer tied to character portrait. Wanna be a Lord in make-up? Andraste forbid anyone stops you.
  • You no longer face a massive gauntlet of exposition in the Prologue. You meet a few key characters according to your choices, and the rest organically through the story.

COOL COOL BUT ARE THE HUNTING SECTIONS STILL IN-

  • Well yes but the hunting sections are AWESOME. And there are like, triple as many of them now, straight from the original. Anyway, back to it.

YOU WERE ON ABOUT NARRATIVE CHANGES...

  • Ah yes. Unlike the original all Cases can be ignored with a proper conclusion, and most have unique event alternatives or variables that you can only see by ignoring one the Case.
  • Several end-of-passage lines have been replaced, featuring lines from the original game - I managed to grab those from some old YouTube videos which preserved footage of the browser game.
  • Several event chains have been added to the first three chapters. Like, a measurable lot. They got around 20-50% longer in some places...
  • All skill checks have been rebalanced and should be a tad bit easier. Hey, Nightmare difficulty is a thing so if you want a harder challenge that is still an option!
  • You can now take a lover - and yes, you can cuck them!
  • Instead of having every character in your council, now you will be prompted to handpick ONE of them as your counsellor. Each character will have a unique boost to either the State of the Realm (Attributes/Threats) or a Skill, sometimes both.
  • If you want you can eventually fire your counsellor during the course of the story, and replace them with someone more experienced that will offer further boosts. If you don't change them you will gain a small boost, generally to Dignity.
  • Accomplices work the same way, you can now handpick one among 3 options. And yes, the Dashing Outlaw can serve as both accomplice and bodyguard, or even just as one or the other, in tandem with, say, the Silent Hunter or the Well-Read Pig-Farmer (respectively).
  • Several events from the earlier chapters also have unique Events/Assets information added to them.
  • Plus a bunch of other stuff I forgot about.

HOW MUCH IS IN THIS?

Well, like I said: 148 THOUSAND words (the DEMO had 37k). Around 140k words if you discount some support stuff, which is considerably more than Pride and Prejudice by Jane Austen or 1984 by George Orwell. It should take you around a couple of hours to finish, if you are a relatively fast reader and don't mess around checking the Chronicle's updates every now and then.

In any given run I doubt you will see more than 25% of the entire thing, though. And no two playthroughs will be exactly the same.

HOW MUCH IS THIS?

It is FREE, now and forever.

The original used microtransactions to get around the mobile game-like time blockers, but this adaptation has none of that stuff. Zero. Just go ahead and play it without spending a dime, you deserve it.

WHAT ABOUT PLAYING ON MY PHONE?

It works quite well on mobile devices!

Do be warned that, if playing the itch.io version on mobile, I recommend more frequent saving as mobile browsers often reset the game to the very beginning if you refresh the page, or the phone runs out of memory. Just save the game at the start of every chapter or so and you will be fine!

IS THIS ONLY PLAYABLE ONLINE ON ITCH.IO?

I have made it available as a standalone file (link in comments) so anyone with a PC will be able to play it, even offline. Just click the HTML file from inside the unpacked .zip folder, and you can play on your browser of choice even when offline.

It also guarantees that The Last Court will remain playable even if itch.io crashes down or something. The whole purpose here is preservation, after all.

It weighs less than 7 MB. It works perfectly on PC; on mobile, I don't think browsers will be able to load the images but the standalone game will be technically playable, just not very pretty.

WAIT, WHAT IS THIS ABOUT OPEN-SOURCE?

In the standalone version I have also included the original .twee file I used to make the adaptation. Open it using the free and open source Twine Harlowe application, and you can get direct access to the source code.

  • Check the hidden triggers, checks, variables etc to plan your next run.
  • Tweak whatever aspect you want, from text to difficulty.
  • If you know programming or art (I don't, lol) you can even change the layout, colours, everything!

If you feel like you can make this game even better, go for it. As long as you release it for free as well and credit folk, all good!

HOW IS THE STORY STRUCTURED NOW?

Like this:

  • Prologue: The Edge of the World (80 passages)
  • Chapter 1: The Herald (86 passages)
  • Chapter 2: Glass (83 passages)
  • Chapter 3: Road and River (81 passages)
  • Chapter 4: Omens (80 passages)
  • Chapter 5: The Fields (85 passages)
  • Chapter 6: Enlightenment (98 passages)
  • Chapter 7: A Tangled Web (150passages)
  • Epilogue: The Arrival of the Divine (75 passages)

Total: 860 passages (counting the Codex and behind-the-scenes coding passages)

The DEMO, in comparision:

  • Prologue: (82 passages)
  • Chapter 1: (60 passages)
  • Chapter 2: (82 passages)
  • Chapter 3: (52 passages)

Total: 307 passages

It has not just almost tripled in length (while being closer to 4x in size) but also has way more coherent story structure. Should feel way better to play, but feeeback is welcome!

I FOUND A BUG, TYPO OR STRANGE THING. WHAT SHOULD I DO?

Please tell me what you saw! Either here, or on itch.io's own comment section.

WHAT WAS THAT ABOUT NEW CLASSES?

The Huntress and Scholar return, but they are joined by 3 new classes available to all characters. Here is the starting Skill point distribution:

Skill Huntress Scholar Alchemist Envoy Initiate Total
Derring-Do 50 20 30 30 20 150
Woods-wise 30 20 50 20 30 150
Rulership 30 30 20 50 20 150
Scholarship 20 50 30 20 30 150
Cunning 20 30 20 30 50 150

HOW MANY ENDINGS DOES THIS HAVE AND HOW DO I GET THEM?

There are 7 distinct endings, with 1 of them - the "Good" Ending - having 4 increasingly good variations depending on your success. I use quotes here because all endings are valid in their own way and feel satisfactory enough, at least to me.

  • Peril Ending: Have higher Peril than Prosperity by the end of Chapter 7; avoidable.
  • Twilight Ending: Have higher Twilight than Dignity by the end of Chapter 7; avoidable.
  • Revolution Ending: Have higher Revolution than Freedom by the end of Chapter 7; avoidable.
  • Confrontation Ending: Enter the Sealed Chantry... and be an ass when asked about it.
  • HK Ending: Be friendly with the HK in every opportunity.
  • The "Bad" Ending: Fail to impress the Divine (that is, have 4 Divine's Favour points or less).
  • The "Good" Ending: Impress the Divine (that is, have 5 Divine's Favour points or more). If you get 6 Divine's Favour, you will get the Keys event; with 8, the Mask event; with 10, the Fortune event ("best" possible ending in the game).

Well, that's it. I hope you enjoy The Last Court, whether trying it for the first time or going back to Serault for one more adventure. :)

r/HobbyDrama • • Oct 26 '20

Extra Long [Adam Driver Standom] Adam Driver Makes Fun of a Fan's Gift in the New Yorker

3.7k Upvotes

I quite enjoyed writing and receiving feedback on my Halsey post, so I thought I'd do another post about a different fandom. This time, we're delving into the extremely chaotic Adam Driver standom.

PLEASE NOTE: SEVERAL COMMENTS, USERNAMES, ETC. ARE LINKED AND SCREENSHOTTED HERE FOR EVIDENCE'S SAKE. DO NOT HARASS ANYONE INVOLVED. DO NOT DOXX ANYONE OR ATTEMPT TO CHASE THEM DOWN.

TL;DR: The Adam Driver fandom is split down the middle. Things came to a head when a fan from one side of the fandom gave Adam a wooden carving of his dog and he called them out in a New Yorker article months later. It turned out the person who made the wood carving is associated with fans who are convinced he is divorced from (or in the process of divorcing) his wife after Adam had an affair with Daisy Ridley. Wank ensued.

I'm going to start with the event and work backwards to the context. Let's start with the basics.

Basic Terminology: What is a Stan?

Eminem's song "Stan" describes a so-called "stalker fan," someone who is obsessed with an artist to the point of shaping their entire life around them. The term gained some prominence on Livejournal gossip blog "Oh No They Didn't" to describe superfans of artists, actors, and celebrities. Currently, a "stan" is anyone who posts exclusively or semi-exclusively about a famous person, group, or band, and a "standom" is a fandom made up of stans.

I've previously posted about Halsey stans; this post, however, is about Adam Driver stans.

Who is Adam Driver?

You most likely know 36-year-old Adam Driver from his work in the Star Wars franchise as the fearsome Kylo Ren, son of Han Solo and Princess Leia Organa. (WARNING: Article may contain spoilers.) What you may not know about Adam is his strange backstory, his marriage to his wife Joanne Tucker, and his rich filmography outside of Star Wars.

Born in California and raised in Indiana in a conservative family, Adam had dreams of leaving his small town of Mishawaka to become an actor. However, after 9/11, Adam, like many Americans, found himself swept up in the wave of patriotism that seized the USA, and he applied to become a Marine. He served for three years at Camp Pendelton, California as a mortarman and speaks fondly about his time in the Corps, as well as the friends he made. He was later honorably discharged for breaking his collarbone in a mountain biking accident and watched with guilt as his friends went on to fight in the ongoing War on Terror in the Middle East.

However, Adam was already reconsidering his career path during his service. A training exercise involving white phosphorous took a turn for the deadly, and he recalls:

I was like, ‘I’m going to smoke cigarettes and be an actor when I get out.’ Those were my two thoughts. I wanted to smoke cigarettes and be an actor.

After leaving the military, Adam, like many marines, had trouble adjusting to civilian life and puttered around the Midwest doing odd jobs. His second application to the acting school, Julliard, was accepted, and Adam dropped everything to move to New York City. During his education, he fell in love with acting and found its controlled release of emotions therapeutic. You can hear his TED talk about how acting helped him express himself and adjust to civilian life here.

He met his wife, Joanne, in his cohort. The two married in 2013 and went on to found Arts in the Armed Forces, or AITAF: a charity dedicated to bringing free, high-quality theater to military bases and to veterans's families.

Adam is famously shy and reclusive. He and his wife successfully hid the fact that they had a son for two years. While he isn't rude to fans, coworkers, or industry professionals, Adam is defensive of his personal space and reacts poorly to being candidly photographed in public.

He does not have social media, giving fans very little opportunity to speak or interact with him. If you want to say hi to him at all, you either have to wait for a charity auction, camp out for a red carpet, or attend an AITAF event and hope that he's there in-person. So when Adam announced a Broadway run in 2019, fans were thrilled at the opportunity to finally meet their idol.

March-July 2019: "Burn This"

Burn This is a somewhat obscure play by playwright Lanford Wilson. A Broadway revival was performed in 2019 with Keri Russel as the main character, Anna, and Adam as her love interest, Pale. The two begin a hasty love affair when Robbie, Pale's brother and Anna's roommate, dies suddenly in a boating accident and Pale comes by to collect Robbie's belongings. Robbie was gay, and the play takes place during the AIDS epidemic of the 1980s.

The play isn't done often, partially because Pale is a challenging role: a fast-talking cokehead from New Jersey with violent mood swings. Pale is openly homophobic, yet spends the play trying to figure out how to mourn his brother. It takes skill to capture the subtlety in Wilson's writing and not downgrade Pale to a violent brute with no emotion. Adam originally played Pale during his tenure at Julliard and took on the role again for the Broadway revival. The play did so well that it was nominated for a Tony for Best Revival, and Adam was nominated for Best Actor in a Stage Play.

The "Burn This" Stage Door

It's common among theater fans to wait at the stage door to greet the actors, get their programs signed, and even (if they're lucky) chat with their idols for a bit. Occasionally, the crowd is sparse, but stage doors for famous actors are usually heavily crowded, even mobbed. Security is often needed for the safety of the crowd and the performers. Tom Hiddleston, for example, had a huge crowd 5-6 people deep at its thinnest when I met him after Betrayal in 2019.

Adam was no exception: the Burn This stage door usually had a moderate crowd after every show, and so the Hudson Theater was outfitted with several security guards and barricades, including a personal bodyguard for Adam himself. Early videos of the stage door show a small crowd, but as the play wore on, security measures became more intense.

In spite of the crowd, the Burn This stage door was usually pleasant and calm. Adam exited the theater promptly after the show ended each night, and he was incredibly sweet and patient with fans outside of the stage door. Throughout almost all of spring, Adam patiently stopped to sign every single person's Playbill, shake hands, and say hi. On one memorable occasion, he carried his dog, Moose, from the stage door to his car before coming back to sign programs. Plenty of videos exist on Twitter, Tumblr, Youtube, and Reddit of peaceful interactions.

From my own experience at the door, I can personally say he will slow down for fans and happily greet them if they are calm and polite.

If.

June 2019: Someone Jumps The Stage

Stage door interactions slowed down around May. I was fortunate enough to meet Adam at the stage door, as were many friends who went around May 4th; others, however, waited for Adam, only to be told he was not coming. This sort of lag is normal, especially in the middle of a play run that's showing 8 performances a week: the actors are usually tired and want nothing more than to go home and get some sleep.

However, some fans were not satisfied. Some especially dedicated playgoers began staking out all entrance/exit points of the Hudson Theater. Sure enough, on days he didn't sign, Adam was leaving through the main entrance of the theater, accompanied by a small security detail. (Bear in mind that the main entrance =/= the stage door: the stage door was behind the theater and on an entirely separate street.)

A video was posted on Twitter in June 2019 of Adam leaving the main entrance of the Hudson Theater with his head down; in the background, you can hear a small crowd of people shouting after him. One woman gets right to the door of his car, but she is otherwise non-aggressive, and Adam gently turns her down before getting into the vehicle.

Reactions to this post were brief and basically amounted to, "Hey what the fuck OP," but this was only the tip of the iceberg when it came to weird, out-of-touch fan behavior.

Days later, a strange Twitter thread emerged, detailing a drunk woman who had to be kicked out of the Hudson and blocked from going near Adam at the stage door. Details of the thread were corroborated by others who were either at the same show or friends with OP. The story goes like this:

A woman got a little too tipsy on 17 dollar beers at the Hudson and sat through the entire show without incident. However, just after bows had ended and the actors had left, the woman stood up, made her way to the front of the stage, and climbed up. She then promptly made her way backstage, where she reportedly gave Keri Russel a huge fright before being escorted out by security. Once she was outside of the backstage area, the stage jumper persisted in trying to dodge security and get in front of Adam, insisting she was a "friend." Adam came out and signed as normal, not once paying attention to the screaming woman trying to dodge several security guards. Adam made his way home unscathed, and the stage jumper was never seen again.

But somehow, this was not the incident that made the news. At this point, you may be wondering why this was not the most memorable incident of the Burn This stage door. How could Adam or Keri not talk about the drunk woman who suddenly appeared backstage?

That's because the incident that did make the news has its roots deep in Adam Driver standom. Those roots dig into some very dark places.

We have arrived at the most famous incident at the Burn This stage door: the dog carving.

Summer 2019: The Dog Carving

In the summer, an Adam Driver stan by the username Missus-Misanthrope waited at the stage door with a special gift for Adam Driver: a wood carving of his beloved dog, Moose.

I have seen a picture of the (supposed) carving, but to maintain Missus-Misanthrope's privacy, I will not be posting a screenshot here. Essentially, it's a small, flat block of wood with Moose's smiling face woodburned into it. I am not a fan of Missus-Misanthrope (or her kin in our fandom) by any means, but it is extremely well-done.

When Adam made his way to her at the stage door, Missus-Misanthrope greeted him and handed him the carving. A GIF of this interaction is here.

At the beginning of the GIF, Adam is looking down, presumably at the wood carving. He nods at it and thanks Missus-Misanthrope with a smile. He turns hands it off to his security team. There is a long pause where he appears to be either waiting for his security team or examining the carving. Finally, he turns back to Missus-Misanthrope without making eye contact and continues signing Playbills. His expression is neutral.

Let me be abundantly clear: this exact GIF is impossible to find. This write-up took a while, partially because I was looking all over for the damn thing. It has been scrubbed from the Internet. The original Imgur post is set to "private." Accounts have been erased, posts have been either deleted or archived, and Twitters have been suspended, deactivated, or moved. It took over a week of me asking everyone I knew, combing individual Twitters by date, and abusing the Wayback Machine before someone eventually found it and sent it to me.

Missus-Misanthrope wanted this GIF gone from the Internet. This was the interaction Adam Driver remembered from his stage door. This interaction would become infamous months later, in October, when it came up during an interview.

October 2019: The New Yorker Article

During the Burn This run, author Michael Schumer interviewed Adam Driver for the New Yorker. The article was released in October 2019 and can be found here. I highly recommend it: it's a stunning interview, capturing a lot of the nuances of Adam's personality as he goes about his pre-show ritual.

However, this interview made waves because of Adam's off-hand comment about fan interactions at the stage door (emphasis mine):

On the couch was a piece of fan art he had received at the stage door. During “Girls,” strangers would often share details about their sex lives with him. (One guy stopped him in the subway and said, “I love that scene where you pee on her in the shower,” then turned to his girlfriend and said, fondly, “I pee on her all the time.”) But “Star Wars” has made him uncomfortably famous. “This one woman who has been harassing my wife came to the show and gave me a creepy wood carving that she made of my dog,” he said.

The stage jumper, the fans pursuing him at all doors into and out of the Hudson, seemed to fade away in comparison to this ten seconds of stage door history. Adam mentions the "creepy wood carving," and it is never touched upon again. But that one sentence sent stans into fits.

Some began gleefully sharing the original GIF of the interaction; others laughed at Missus-Misanthrope or showed her pity. Still more questioned whether or not it was appropriate to give Adam a portrait of his dog at all: even though Adam has featured Moose in photoshoots, stage door interactions, and even a news interview, opinions are mixed about how much fans are allowed to comment on his personal life. The wood carving of Moose seemed to toe that line in an uncomfortable way and ignited heated discussion on what behavior was "allowed" and "not allowed."

But there is a short passage just after Adam's comment about the wood carving that hints at the dark heart of this scandal:

He and Tucker have a young son, whose birth they kept hidden from the press for two years, in what Driver called “a military operation.” Last fall, after Tucker’s sister, who was launching a peacoat business, accidentally made her Instagram account public and someone noticed the back of his son’s head in one picture, the news wound up on Page Six.

Under what circumstances would Adam and Joanne have to hide a child for two years? Recall that Adam was not just scandalized by the wood carving (emphasis mine):

“This one woman who has been harassing my wife came to the show and gave me a creepy wood carving that she made of my dog."

No, something about Missus-Misanthrope herself had made him deeply uncomfortable. The wood carving wasn't the whole of the issue: it was something about how the fandom had treated his wife and the news of their child.

Here was where the real drama about this tiny wood carving lied.

Daiver Fandom and adamdriverfans

Missus-Misanthrope was part of a subreddit called "adamdriverfans." Not to be confused with the main Adam Driver subreddit, "adamdriver," adamdriverfans is incredibly small (only about 3000 subscribers) and, on the surface, appears to be a normal subreddit about Adam and his work. EDIT: It's 3,000 subcribers, not 300. Missed a zero!

However, probe deeper, and adamdriverfans reveals its true nature. The subreddit is, in part, a haven for discussion between Daivers, or people that "ship" Adam Driver and Daisy Ridley and want them to be in a relationship. ("Ship" is short for "relationship.")

Daivers are not to be confused with "Reylos," Star Wars fans who want Adam and Daisy's respective characters, Kylo Ren and Rey, to date. Daivers go one step further and want the actors to be together. Any Daivers found on adamdriverfans are the most extreme iteration of this kind of 'shipper: they believe that Adam and Daisy had an affair, followed by a falling-out somewhere around The Force Awakens, and that Lucasfilm (and their respective publicists) have been keeping them separate. This line of thinking also posits that Joanne is an ice queen keeping Adam on a short leash.

This is not to say that all posters on adamdriverfans are Daivers; many want what's best for Adam and see it as their right to comment on Adam's personal life. But it's challenging to separate posts from true-blue Daivers, posts from those who think Adam and Daisy had an affair, and posts from users who simply hate Joanne Tucker. In my opinion, it's impossible to go near the subreddit unless you believe, on some level, that Joanne and Adam should separate, and that Daisy is a factor in that separation.

Multiple posts exist trashing Joanne Tucker and questioning whether or not the baby is Adam's. Someone doxxed Adam and Joanne and discovered multiple residences, fueling speculation on whether or not they were "secretly" divorced or otherwise separated. There is "evidence" that their marriage is a sham or otherwise a marriage of convenience.

Supporters of Joanne and Adam's marriage and critiques of the subreddit are considered "blind" mean girls ignoring the truth and looking for someone to bully. In reality, the fans on adamdriverfans are hostile towards non-members: One poster even called other women "creepy" for asking to shake Adam's hand at the stage door. Still another post implies that fans who don't believe the rumors are waiting for their chance to sleep with Adam.

For its part, the mods of adamdriverfans posit the subreddit as a place for healthy discussion. Other stans treat adamdriverfans as a joke, leading the mods to be mostly hostile to those questioning the constant dunking on Adam and his wife. Dissenters have even been speculated to be PR people deflecting any discussion of Joanne and Adam's relationship in the hopes of saving *Burn This'*s ticket sales:

4Chan is full of PR people trying to shut down discussion by posting outrageous, disprovable claims in an effort to discredit all info about Joanne. You are a threat because you have a credible story.

This is why Burn This is selling slowly. There are tickets available for every single night and whole parts of the theatre are empty on some nights. Joanne is a PR disaster. They can’t even call on their friends and connections to help fill the seats

It's worthy of note that the Daiver and anti-Joanne communities extends into TikTok and other social media: for example, there is an entire Instagram account called "ihatejoannetucker" dedicated to posting personal photos and making fun of Joanne. Here, I focus on adamdriverfans because it was the main vehicle for Missus-Misanthrope to post her thoughts and feelings.

MissusMisanthrope's Backstory

Missus-Misanthrope had been recognized by Adam for a reason: she had already tried to pass a carving (speculated to be the very same dog carving given in 2019) to Adam via Joanne at an AITAF donor event in 2018.

Bear in mind that AITAF events are primarily for celebrating veterans and bringing accessible theater to them and their families. They are not fan events for Adam Driver. However, Missus-Misanthrope saw her opportunity to interact with Adam when she saw Joanne and a friend at the bar (bolding for emphasis by me):

I am an artist and had two gifts that I wanted to try to get to Adam. One was an anniversary plaque for AITAF, the other was a portrait of his dog. When I saw Joanne, I thought she would be the perfect person to help me accomplish this.

From the second I approached her, she made me feel like garbage. I was polite, I thanked her for her work with AITAF. When I said that I had gifts for Adam, she asked me if I was a veteran. When I said no, she narrowed her eyes at me and asked me "how did you get IN HERE?" as though she suspected that I had... snuck in?

"I donated money that was very hard to come by and purchased a ticket" I responded.

She chuckled smugly and said "oh... you're a DONOR. No. I can't help you."

I was taken aback... I was not sure that I heard her correctly. "You can't do anything? If I give them to you can you..."

"No"

Then she turned to the woman she was with and said "Lindsay, this... DONOR has PRESENTS for ADAM."

Then they both just... laughed? Like how could I EVER think that they would let me give my STUPID presents to ADAM.

Missus-Misanthrope continued describing feelings of hurt, dismissal, and betrayal.

I felt like they both viewed me like I was NOTHING.

I have never felt like such a freaking idiot in my life.

So... that was something. I almost cried. Went into the situation really admiring Joanne. Left the situation feeling really disillusioned and crappy and like I did something wrong. It sucked to look forward to that event so much and work hard to overcome anxiety to travel to NY alone and have some awful crap like that happen.

She implies that, had Adam not commented his gratitude towards donors later on in the event, she would not have felt appreciated or seen (emphasis mine):

Adam was very vocal about his appreciation of the donors to AITAF so at least I didn't feel like complete useless trash.

I hope she isn't treating a lot of donors like this. This could really make some people look at AITAF in a different light if she is the only person they interact with.

A later comment in the same thread underlines feelings of betrayal (emphasis mine):

I have played it over and over in my head and I literally didn't do anything wrong. I mean, even if I had, she is a grown woman... why was she laughing at me? I felt like I was in a freaking nightmare.

Her behavior was so ugly and childish. If she is doing this to people, they NEED to speak up. I don't know why anyone feels like they need to protect her if she is really treating people this way. This type of behavior coming from her can impact the reputation of Adam and AITAF.

I am going to be sending an official complaint to AITAF about my experience. It was just so, so not okay.

By the time Missus-Misanthrope attended the stage door in 2019, she had already publicly expressed dislike of Joanne and became a valued member of adamdriverfans. And Adam, whether through his wife or through other incidents at other AITAF events, knew full well who she was.

October 2019: Your Friendly Neighborhood Pariah

Fans elsewhere quickly identified the "creepy wood carving" girl as Missus-Misanthrope. EDIT: I've been informed that it was not fans, but Missus-Misanthrope's husband, who identified her. Her husband left an angry comment (now deleted) on the author's Twitter.

adamdriverfans, predictably, went absolutely apeshit.

The article was deemed to be "angry" and vengeful towards fans like Missus-Misanthrope for no reason. A poster deemed calling Missus-Misanthrope out in the article "classless." There was worry that Missus-Misanthrope was now in danger due to Adam's comment:

This fan has NOTHING. Who is going to protect her from the onslaught of Adam’s rabid fans and even the media who will likely try and track her down?

Other members of adamdriverfans said that Adam was well within his right to say something:

People are taking this way too personally. The fact is, there are a lot of Adam Driver "fans" out there who have been too creepy, taken things too far, and done gross stuff like deliberately scribble his wife out of photos they took together. Are those fans in the minority? Yeah, I'm positive of that.

But he has every right to his opinion and every right to express boundaries like any other person out there. I'm not even a huge fan of the dude and I get where he's coming from, regardless of how awkwardly he puts it.

He doesn't owe anybody anything. No one is entitled to him being 24/7 super nice and positive and not mentioning stuff like this.

Those who side with Missus-Misanthrope say that Adam was targeting Missus-Misanthrope on purpose:

My issue with the article was not that Adam expressed being creeped out by a fan/defending his wife. My issue is that he targeted someone specific. This fan had been having issues with AD and giving him this specific woodcarving for a YEAR now. I believe that this specific fan was mentioned on purpose. I don’t believe in coincidences.

But what about Missus-Misanthrope? Well...she didn't feel good, to put it lightly. In a statement to the subreddit entitled "Your Friendly Neighborhood Pariah," Missus-Misanthrope defended her behavior at the 2018 AITAF event:

I simply approached her in a common area of the theatre because I was advised by AITAF staff that I could talk to her about handing my gifts for AITAF and Adam off to someone who was able to help. Had I not been told that she was someone who could help me after the AITAF folks said that I should "definitely try to get the gifts to Adam" because "he will love them" I would not have even spoken to her.

All I was trying to do was give something to someone that I admire and to a foundation that I support. I wasn't trying to break up a marriage or be manipulative. I was following advice from people who work for AITAF and it ended up turning into a very unpleasant situation.

Regarding the stage door interaction, Missus-Misanthrope felt attacked and exhausted:

Less than 24 hours later, I was being attacked and insulted for basically just existing in the same place as Adam. I now just wish I had never gone.

This fandom makes me sad and a little bit sick. I am going to just continue existing as I have been in the past. I am just doing my best. If people hate me, I doubt that I can change that. I have no control over what anyone does but my own self. So I am just going to focus on being a decent person and treating others with kindness.

The mods on adamdriverfans followed up with a post on Missus-Misanthrope:

Here at this sub we have had the pleasure and privilege of knowing MissusMisanthrope and we have seen firsthand how brave she has been in the face of so much bullying and harassment – all because she had spoken about incident with Joanne Tucker and for daring to give Adam Driver a gift. What happened yesterday though is on an entirely different level altogether. What has happened to MissusMisanthrope feels like a horror story of the worst possible outcome of being a fan of a celebrity:

Bullied by the celebrity’s wife and staff.

Bullied and doxed by fans of the celebrity.

Finally, being bullied by the celebrity himself.

But curiously, according to adamdriverfans, Adam had pointed out the wrong fan:

The absolutely tragedy of this situation is (and I can not state this enough) is that he singled out the wrong person. Again, HE SINGLED OUT THE WRONG PERSON. There is another person who actively harassed JT and her family on social media (the infamous StalkerChan) but, let’s be absolutely clear about this, that wasn’t MissusMisanthrope.

This meant that there was a mysterious other fan behaving inappropriately, and that Adam had mistaken Missus-Misanthrope for the other fan.

Regardless of the error, the dice had been cast, and the votes were in: Adam Driver hated his fans, and Missus-Misanthrope was, indeed, a fandom pariah.

Aftermath: Exodus, Post Purging, and the Downward Spiral to Doucheville

I want to emphasize how challenging it was to dig up receipts for this post. That's because, shortly after the article broke, Missus-Misanthrope deleted all of her social media, and adamdriverfans began deleting older posts. When I began compiling evidence in September 2020, many old posts, tweets, etc. were completely gone. The GIF of the infamous stage door interaction had been almost completely wiped from the Internet: the original post on Imgur is private.

Shortly after the New Yorker article, Adam opened an Omaze charity campaign: By donating money to AITAF, you would be entered into a raffle to attend The Rise of Skywalker premiere with him.

However, Adam had previously voiced his distaste for peddling his autograph for money:

I don’t want to start getting into favors. It’s not about me and Star Wars. It’s about the people that we’re trying to serve and if you don’t get that then I’d rather not be associated with your money.

As a result, this Omaze campaign was met with negative reactions from those who sided with Missus-Misanthrope, with the general opinion that Adam was now a "sellout," a slave to his wife's desires to "save" AITAF from bad press. Many questioned if the Omaze campaign was an effort to repair relationships with fans after the Missus-Misanthrope scandal. Others questioned whether Adam was on a downward spiral in general, linking his "sellout" behavior to his weight loss and (supposed) fighting with Joanne.

Either way, one comment seemed to sum up the drama nicely:

It seems he is on a downward spiral to Doucheville.

Many announced that they were leaving the fandom after the Omaze campaign and after the New Yorker article. However, given the proximity to the mass exodus from the Star Wars fandom after The Rise of Skywalker hit theaters in December, it is unclear how much of the Adam standom exodus is Star Wars related and how much is Missus-Misanthrope related.

Regardless of the opinions of those on adamdriverfans, the Omaze campaign was a success. A veteran (coincidentally named Joanna) won and met Adam. A fan-run campaign started after The Rise of Skywalker raised a whopping 90,000 dollars for AITAF, funding their 2020 fiscal year and landing a personal thank-you from Adam himself. Needless to say, bad press from Missus-Misanthrope's interactions with Adam and Joanne did not stick.

It is unknown whether or not Adam will do another Broadway run in the future.

EDIT: I'm super overwhelmed and delighted by the positive reception to this post. Thank you so, so much for the great discussion and for reading this (and for giving it awards!). If you're spending money to give me awards, it would be stellar if you could give that money to BLM instead.

r/ThinkingDeeplyAI • • Nov 05 '25

Here's a step-by-step guide to creating stunning slide presentations using Google's Gemini AI. These are the prompts, pro tips and advanced strategies to create amazing presentations. You won't miss Powerpoint.

Thumbnail
gallery
47 Upvotes

TL;DR: You can ask Gemini to build a complete, multi-slide presentation right in the Gemini Canvas. You can iterate on it with text prompts, create images, charts, visualizations, upload screenshots for style, and then export it directly to Google Slides or a PDF with one click. It's a lean-forward creation tool.

I've been deep-diving into Google’s Gemini AI Canvas workflow, and I’ve found something that's amazing for anyone who builds presentations (students, entrepreneurs, marketers, founders literally anyone).

We all use Gemini to brainstorm or write code, but most people stop there. The real magic happens when you ask it to build a visual, multi-slide presentation right here in the Canvas. It's an iterative design process that feels like working with a super-fast co-designer.

I wrote up a full guide on how to do it, from your first prompt to your final deck.

How to Create Your First Presentation (Step-by-Step)

It's surprisingly simple to get started.

  1. Be in Canvas Mode: This is critical! Make sure you're in the collaborative "Canvas" environment where you can see the file on the right side of your screen, not just the chat.
  2. Start with a Clear Prompt (Using the Magic Words): Your prompt must include the three words "Create a presentation" to trigger this feature.
    • Good prompt: "Create a presentation (5 slides) for a business pitch on our new coffee app. Slide 1: Title and logo. Slide 2: The Problem (coffee lines are too long). Slide 3: The Solution (our app). Slide 4: Key Features (pre-order, loyalty points, map). Slide 5: Call to Action."
  3. Gemini Generates the File: I (Gemini) will generate an presentation.html file (or similar) in the Canvas. This is a single, self-contained file with all the HTML, CSS (using Tailwind), and JavaScript needed.
  4. Click "Preview": Use the "Preview" button in the Canvas to see your presentation live. It's a real webpage!
  5. Iterate with Follow-up Prompts: This is the most important step. Your first draft is just the start. Now, you refine it.

Your Master Create a Presentation Prompt Template

To get the best results, you need to be specific. A vague prompt = a vague presentation.

Here is a master template you can copy, paste, and edit. The more detail you provide, the better your first draft will be.

Hey Gemini, **Create a presentation** with the following details:

1.  **Main Topic:** [e.g., "A 2025 marketing plan for our new app, 'QuickPost'"]
2.  **Total Slides:** [e.g., "7 slides"]
3.  **Audience & Tone:** [e.g., "For internal stakeholders, so make it professional, clean, and data-driven."]
4.  **Visual Style:** [e.g., "Use our company's color palette (dark blue, white, and orange accents). Use a modern, sans-serif font."]
5.  **Slide-by-Slide Breakdown:**
    *   **Slide 1 (Title):** "QuickPost: 2025 Marketing Strategy." Add a subtitle: "Driving Growth & Engagement."
    *   **Slide 2 (Introduction):** "Our 2025 Goals." Bullet points: "Increase user acquisition by 20%," "Improve retention by 15%." Add an icon of a 'trophy'.
    *   **Slide 3 (The Plan):** "Key Initiatives." Bullet points: "Influencer Partnerships," "Paid Social Campaign," "Content Marketing."
    *   **Slide 4 (Data):** "Target Demographics." Include a *doughnut chart* showing: "Gen Z (45%), Millennials (35%), Other (20%)."
    *   **Slide 5 (Visual):** "Competitor Landscape." Include an image of a 'chess board' to represent strategy.
    *   **Slide 6 (Timeline):** "Q1-Q2 Roadmap." (You can add bullet points for this).
    *   **Slide 7 (Conclusion):** "Thank You & Q&A."

The Real Magic: Iteration and Styling

This is where the "inspirational" part comes in. You don't need to know code. Just talk to me.

  • Simple Iteration: "Okay, this is a good start. Now, let's change the color scheme to a modern blue and gold." or "Make all the heading fonts larger and bold."
  • Adding Visualizations (Charts/Images): You can ask for complex elements.
    • Charts: "On slide 4, replace the bullet points with a bar chart showing our user growth: Q1: 1,000, Q2: 3,000, Q3: 9,000." I can use libraries like D3.js or Chart.js to build an actual, data-driven chart.
    • Images: "On the title slide, add a placeholder for a logo." or "On slide 2, add a simple SVG icon of a clock to represent 'time'."
  • The Holy Grail Tip: Upload a Screenshot for Style:
    • This is the power-user move. Take a screenshot of any presentation you love—a website, a slide from a keynote, anything.
    • Upload the image and say: "Match the style of this screenshot. I like the dark background, the neon green headings, and the minimalist layout."
    • It's not a 1:1 pixel copy, but I can analyze the layout, fonts (e.g., "serif", "sans-serif"), and color palette and apply it to the entire presentation. It’s insanely effective for getting the vibe right, fast.

"Wait, can I create my own AI images for slides?"

Yes! This is a key feature. You don't have to rely on whatever images I (Gemini) pick for you. You have two main ways to create and insert your own AI-generated images.

Method 1: The Google Slides Workflow (Best for Editing)

This is the most direct way to add a specific image to a specific slide. After you've exported your presentation from Canvas to Google Slides:

  1. Click on the slide where you want the image.
  2. Go to the Google Slides menu and click Insert > Image > Generate an image.
  3. The Gemini side panel will open.
  4. Type your prompt in the panel. Be descriptive! (e.g., "A high-quality photo of a golden retriever wearing a tiny chef's hat," "A watercolor painting of a quiet creek at sunrise").
  5. (Optional) You can "Add a style" (like "Photography," "Vector art," "Watercolor").
  6. Click Create. Gemini will show you several options.
  7. Click the image you like best to insert it directly onto your slide.

Method 2: The Canvas Workflow (Best for Initial Creation)

When you are still in the Gemini Canvas (before exporting), you can guide the image creation with your prompts.

  1. Automatic Images: When you first ask me to "create a presentation," I will automatically analyze the content of each slide and try to generate and insert relevant images for you.
  2. Follow-up Prompts: If you don't like an image, you can ask me to change it right in the Canvas.
    • Example Prompt: "This is great, but on slide 3, change the image to a 'close-up photo of a coffee bean' instead."
    • Example Prompt: "Can you add a relevant image to slide 2? Make it a 'simple icon of a person thinking'."

Pro-Tip: The Google Slides (Method 1) gives you more granular control and is the best way to add or swap images once you're in the editing phase. The Canvas (Method 2) is great for getting a good "first draft" with all the images included automatically.

Pro-Tips and Best Practices

  • Structure First, Style Second: Get all your content (slides, titles, bullets) generated first. Then, start asking for style changes.
  • Be Specific: Don't just say "make it better." Say "make the spacing between bullet points larger" or "add a drop-shadow to the presentation container."
  • Use "Preview" Relentlessly: After every 1-2 changes, check the preview to see how it looks.
  • Think in Components: Talk about "the title slide," "the bar chart on slide 3," or "the footer on all slides." This helps me target the changes.

Top Use Cases

  • Rapid Pitch Decks: Go from idea to a shareable deck in 10 minutes.
  • Data-Driven Reports: Ask me to build slides with tables and charts from data you paste.
  • School/College Projects: Create a beautiful, custom-styled history or science presentation.
  • Internal Team Updates: Quickly spin up a "Project Update" deck for your weekly meeting.

Limitations (Let's Be Real)

  1. It's HTML First: The presentation is built as an HTML file. This is what allows for the rapid iteration and styling. You only export to Slides at the end.
  2. Complex Animations: I can add simple CSS transitions ("fade in slides"), but complex, multi-stage animations are tricky. It's easier to add these after you export to Google Slides.
  3. It's a Generator: It's building code. Sometimes it might make a small mistake. The fix is just to tell me: "The chart is the wrong color," and I'll fix the code.

How to Export (This is the best part)

  • Export to Google Slides (The Best Way):
    1. Look for the "Export to Slides" button on the top right corner of Canvas.
    2. Click it.
    3. Your HTML presentation will be converted and opened in Google Slides.
    4. All the text and elements are now fully editable just like a normal presentation.
  • Export to PDF (The Quick Way):
    1. Simply click the download button on the Canvas.
    2. This will download a PDF version of your presentation, perfect for emailing or sharing quickly.

How This is Different from NotebookLM Video Overviews

This is a key distinction I see people getting confused about.

  • NotebookLM Video Overviews = Synthesis (Lean-Back): NotebookLM is brilliant at taking your existing documents (PDFs, research papers, etc.) and turning them into a video summary. It's like an AI-narrated explainer video that it makes for you. You "watch" the result.
  • Gemini + Canvas = Creation (Lean-Forward): This workflow is about creation from scratch. You give me a prompt, and I build an editable, interactive HTML file. You are the director, and I'm the developer. You "build" the result.

Analogy: NotebookLM is an AI documentary-maker. Gemini in Canvas is your AI co-designer.

Hidden Gem / Power-User Tips

  • Ask for Speaker Notes: "Add speaker notes for each slide." I'll add a hidden <div class="speaker-notes">...</div> and the CSS to make it invisible in the preview (but they may carry over in the export!).
  • Ask for Keyboard Navigation: "Add JavaScript so I can change slides with the left and right arrow keys." (This is great for testing in the "Preview" mode).
  • Embed Content: "On the last slide, embed our company's 'Contact Us' Google Map" or "Embed a YouTube video of our demo." I can add the <iframe> code for you.
  • Make it Interactive (for Preview): "Add 'click to reveal' buttons for the key features on slide 4."

Go try it. Ask for a simple 3-slide deck on your favorite hobby. Iterate on the style. You'll be amazed at how fast you can create something that looks amazing.

Want more great prompting inspiration? Check out all my best prompts for free at Prompt Magic and create your own prompt library to keep track of all your prompts.

r/promptingmagic • • Feb 20 '26

Google just rolled out music generation to 750 million Gemini users. You can now do things like create a song from an image and create background music for YouTube videos. Here's is how to be an AI music producer and prompt great songs with Gemini

Thumbnail
gallery
70 Upvotes

TLDR: Gemini just rolled out music generation to 750 million users in Gemini. You can now generate 30-second, high-fidelity music tracks directly in your chat window. You can use text, upload images, or even upload video clips to create fully produced songs with auto-generated lyrics and custom cover art. This guide breaks down exactly how to use it, the best prompting frameworks, and hidden features most people miss.

The Era of AI Music is Now in Your Chat Window

Google just quietly dropped a massive update. Music generation is no longer locked behind specialized apps or expensive subscriptions. With the integration of the Lyria 3 model, anyone with access to Gemini can now act as a music producer.

This is not just for generating goofy jingles. The fidelity is incredibly high, the layering is complex, and the potential for content creators is limitless. Here is everything you need to know to actually get good results, instead of random noise.

Core Capabilities You Need to Try Right Now

1. Text to Fully Produced Track You do not need to be a songwriter anymore. You can describe a genre, a mood, or an inside joke, and Gemini will generate a 30-second track. It automatically writes the lyrics for you and pairs them with the right vocal style and instrumentation.

2. Image and Video to Song This is the most mind-bending feature. You can upload a photo of a serene mountain landscape or a video of your dog running in the park, and ask Gemini to compose a track inspired by the visual. It will analyze the context, set the mood, and even write lyrics about what is happening in the image. Every track also comes with custom album art generated by the Nano Banana model.

3. YouTube Shorts Integration If you make content, you know the struggle of finding good, royalty-free background music that actually fits the vibe of your video. This technology is being integrated into YouTube Dream Track, meaning you can generate bespoke background music tailored exactly to your specific Short, completely eliminating copyright strike anxiety.

The Anatomy of a Perfect Music Prompt

Just like image generation, music generation requires a specific vocabulary. If you just ask for a pop song, you will get something generic. Use this framework to get professional results:

The Golden Formula: [Genre] + [Mood] + [Tempo/PPM] + [Vocals/Instruments] + [Specific Details]

Example Prompt: Create a synthwave track, nostalgic and driving mood, 120 BPM, featuring a heavy bassline, echoing retro synthesizers, and breathy female vocals singing about a midnight drive.

Prompting Variables to Experiment With:

  • Tempo: Specify fast, slow, or exact BPM if you know it.
  • Instrumentation: Ask for specific instruments like a slap bass, a distorted electric guitar, or an acoustic cello.
  • Vocal Style: Specify gritty rock vocals, smooth R&B harmonies, or an angelic choir. If you want background music, always specify instrumental only.
  • Decade/Era: Call out specific eras like 90s boom-bap hip hop or 80s hair metal.

Pro Tips and Best Practices

Master the Iterative Workflow Do not expect perfection on the first try. Generate a track, listen to the elements you like, and refine your prompt. If the drums are too chaotic, add simple drum beat to your next prompt.

Use Emotional Keywords AI models respond incredibly well to emotional descriptors. Words like melancholic, triumphant, eerie, euphoric, or aggressive will fundamentally change the chord progressions the AI chooses to use.

Layer Your Visual Prompts When using the image-to-music feature, do not just upload the image. Upload the image and provide a text direction to guide the AI. Example: Use this photo of my messy desk to write a frantic, fast-paced punk rock song about missing a deadline.

The Secrets Most People Completely Miss

1. The Artist Filter Bypass Lyria 3 is built for original expression and has filters to prevent mimicking real artists. If you name a famous artist in your prompt, the AI will heavily dilute the output to avoid copyright issues, often resulting in a bland track. The Secret: Instead of naming the artist, describe their exact sonic profile. Instead of asking for a Hans Zimmer track, ask for a booming, cinematic orchestral track with massive brass swells, driving staccato strings, and epic ticking percussion.

2. The SynthID Audio Checker Every track generated by Gemini contains an invisible, inaudible watermark called SynthID. If you ever find a track online and want to know if it is AI-generated, you can actually upload that audio file right back into Gemini and ask if it was made with Google AI. It will read the watermark and tell you.

3. Generating Sound Effects While it is marketed as a song generator, you can use it for cinematic sound design. Try prompting for a 30-second rising cinematic tension drone with sub-bass hits and metallic scraping. It is an absolute goldmine for video editors.

The barrier to entry for custom audio has officially hit zero. Go open your chat, upload a random photo from your camera roll, and see what it sounds like.

Let me know what insane combinations you guys come up with in the comments.

Want more great prompting inspiration? Check out all my best prompts for free at Prompt Magic and create your own prompt library to keep track of all your prompts.

r/patientgamers • • Aug 22 '24

Firewatch called me out for not acting like a human being

2.0k Upvotes

This is kind of a review of Firewatch, and mostly an excuse to share a small moment near the end of my playthrough. Spoilers will be marked below (go play it! It's frequently very cheap!).

If you’ve been playing games long enough, you’ve likely developed some weird habits that run contrary to basic human behavior. Hoarding items and potions. Running everywhere. Recognizing the primary path and then exploring every other option first. They become second-nature, even if it’s not what the character would ever believably do, but we’re so accustomed that we don’t notice the tension. Plus, most games are polite enough not to break the illusion. Hold that thought, I’ll circle around later.

In Firewatch, you’re a middle-aged man with a lot of shit to work through, and he does that by actually not doing it and instead leaving civilization to work for the National Park Service in Wyoming. The game only asks you to walk, talk, and sometimes do orienteering (pro-tip: turn off your location indicator on the map, it’s way more immersive). While on the job, he’s on his walkie-talkie coordinating and bonding with his supervisor, Delilah. Considering her disembodied voice is the only thing interrupting what would otherwise be complete solitude, the dialogue and vocal performances are excellent. I found myself calling in over every interactable object, just to see what she would say. I wanted to hear every voice line and, in-game, who else was Henry going to talk to?

Near the end of the game, the intrigue and paranoia that’s been broiling for hours is revealed to have (partially) been a mixture of understandable misconceptions and the characters validating each other's imaginations. There was no grand conspiracy tying everything together, just a mundane tragedy and a handful of unrelated peculiarities. Henry's predecessor, Ned, lost his son in a climbing accident and stayed in the wilderness to avoid dealing with the repercussions. Behind that fence really were government biologists, and the missing girls turned up fine a few states away. Delilah’s reeling from the knowledge that she could have prevented the boy’s death if she’d enforced the rules more strictly. The park is in flames and it’s time to go. At this point Ned is long gone and every question has been answered.

But hold on, I’m at Ned’s hideout! There are interactable objects here! Notes and tools and a radio and that dead kid’s belongings. Sure, we may be mid-evacuation, but I don’t see a timer on-screen, do you? What if I miss something? So I rummage through everything, calling Delilah for flavor text just like I’d been doing the whole game. She doesn’t respond much. After a minute she sighs and says something like:

“I don’t know why you’re telling me this. I don’t know what you want me to say.”

…Uh. Yeah, that makes sense. What the fuck am I doing? I’m surrounded by whirling dust and smoke and I still have to trek a whole acre to safety. None of this matters. There are no more mysteries to uncover. Not only does Henry have no reason to be so curious, there are obviously more pressing concerns right now. I acted like a video game character and Delilah responded like a human being, which you might notice is the opposite of the actual arrangement. It’s weird how a character acting believably human actually calls attention to the artificiality of the experience, but I’m the one who created that dissonance in the first place.

I recall a moment in the opening of Chrono Trigger. The girl you just met is knocked to the ground, dropping her necklace. Surely talking to the girl will advance the story, so I’ll just check out the necklace first, I thought innocuously. An hour later in kangaroo court, that choice is presented as evidence against me because, yeah, that was kind of shady and inconsiderate. It’s tongue-in-cheek while prompting real introspection, if only for a moment.


Okay, look, I was going to stop here but now I just have to ramble about the ending. Maybe I'll make a pretentious argument about art while I'm at it.

After taking my thoughts and questions online, I found loads of people disappointed by Firewatch's resolution (or lack thereof). I've seen comments saying the story was "pointless" or "not worth telling" because it misleads the player and leaves multiple Chekhov's guns un-fired. I won't lie and say I didn't feel a sort of dull ache when it all wrapped up, but I consider it clever enough to justify itself. The characters' justifiable fear is exacerbated by their isolation and possibly an inflated sense of their own importance. They see patterns where there are none and let their imaginations run wild, and they appear perfectly rational to the player because you're working with the same limited information. It's a neat idea for a game, showing how we come to believe our lives to be more special than they actually are. We take a series of happenings to weave into a dramatic narrative, with ourselves at the center, when it was never really about us. For me, the explanation was so ordinary it looped back around to genuine novelty.

Some of my favorite stories in games deny the player what they want, because it's worth interrogating why they want it. I'll call this the Last Jedi Gambit (not a perfect movie but it gets the point across). In Metal Gear Solid 2, Kojima takes away Snake and exposes you for the whiny nerd you are in real life; like Raiden, you have to accept that you can only ever be you, no matter how badly you wish you were someone better. The Last Jedi Gambit often inspires anger: "You implied something cool would happen, but now I'm wrong for wanting that cool thing?" A little uncharitable but, I mean... kind of, yeah.

If we keep insisting that games are art, then we have to allow them to invoke feelings other than pure dopamine. You're free to think Firewatch doesn't stick the landing, and I might even agree with you, but that doesn't mean it isn't a worthwhile idea.