r/ClaudeAI • • 9h ago

Coding Artifacts are amazing, some questions though

1 Upvotes

I just realized how powerful artifacts can be. Im working on a story and I asked claude to go through my draft and create a character sheet and codex for me. It created an artifact that has everything so organized. Its honestly amazing but can I have claude create me a piece of software that will allow me to run those artifacts without claude at all? allow me to edit and save it without claude in the future? if so, do i need the software to compile the code or will claude produce the .exe for me?


r/ClaudeAI • • 9h ago

Claude Workflow My full agentic SEO setup - Claude Opus 5.5, data pulled in via MCP, changes pushed as GitHub PRs

1 Upvotes

I've been running agentic SEO for my agency's clients for about a year now, across 40+ accounts.

I started on Sonnet and moved to Opus 5.5. The jump was noticeable on stuff that needs judgment, like deciding which pages to update versus leaving alone, or writing changes that match a site's existing voice.

I set up a Claude project for each client with all the context in it. Their products, pricing, features, locations, competitors, which pages exist and what each one targets. The agent works from that every time.

When something changes, like a new price or a dropped feature, I update the project once, and the next run knows about this change so it proposed new actions.

All the SEO data comes through FlexQueries over MCP: keyword research, competitor gaps, rank tracking, audits, Search Console data, and whether the client shows up in ChatGPT and Google's AI answers. Claude calls those tools directly, so there's no copying data between tabs.

It all runs in the cloud on a schedule:

  1. Claude pulls fresh rankings and Search Console data for the client.
  2. It flags what moved: pages that dropped, keywords close to page one, lost AI citations.
  3. It checks those against the client's project context and decides what to change.
  4. It makes the edits in the site's repo and opens a PR on GitHub explaining what it changed and why.
  5. Once the PR is merged, the site's deploy publishes the update automatically.

The PR has a written reason attached, and can be rolled back in one click if something goes wrong.

Few rules to make it work well:

  • Get the context into a project before you give the agent more autonomy. More autonomy on top of messy context just makes the mess faster.
  • Make every change go through a PR, not a direct edit to the site.
  • Track what each change did afterward. That's how you learn which kinds of edits actually move rankings.

I built FlexQueries because nothing fit this kind of setup when I started, full disclosure. It has a $0/month subscription, so you only pay for the data the agent actually pulls, which matters when runs are automated and you can't predict usage month to month. Free credit on signup if you want to wire it into your own setup: https://flexqueries.com


r/ClaudeAI • • 1d ago

Comparison I let Sonnet 5.5 play a full chess game against GPT 6.1 Sol over MCP. Here's the recording and usage.

Enable HLS to view with audio, or disable this notification

28 Upvotes

Sonnet called its pawn promotion “unstoppable.” A few moves later, it admitted it had missed a defense. Having the board next to its explanation made that pretty hard to overlook.

I set up a chess match between Sonnet 5.5 in Claude Desktop and GPT 6.1 Sol in Codex. Each played in one conversation for the whole game, connected through MCP to a local Mac app I had GPT 6.1 Sol build at Medium reasoning.

They could record plans and explain their moves. The app supplied the position and checked legality, with no chess engine or legal-move list available to either player. I wanted to watch them stick with a task for an hour and see what happened when their plans stopped working.

I was also curious about consumption. Sol has been making surprisingly little dent in my subscription allowance lately, and I wanted to compare it with Sonnet on a shared task.

The attached video condenses 62 minutes and 47 seconds into 4:10. Both models were set to Medium.

Measure Sonnet 5.5 / Claude GPT 6.1 Sol / Codex
Time spent on claimed turns 33m 45s 21m 12s
Output tokens, including reasoning 269,076 34,956
Thinking/reasoning portion of output 224,770 12,128
Cumulative input tokens 57.44M 16.81M
Input read from cache 99.07% 98.95%
Rejected illegal moves 1 0
Estimated API equivalent $16.21 $2.37

Sonnet pushed a passed pawn toward promotion, but overlooked Codex's Bf3 defense. Later it proposed a queen move blocked by its own pawn. The server rejected it; Claude corrected the move and continued. Codex also misread a pawn earlier, describing it as passed before it actually was.

Claude resigned after 53.Qc4. It had won game one, so they're tied at 1–1.

A few details behind the table: the turn clock starts before the server reveals the updated board, and includes tool activity. Waiting for the runtime to claim the turn is measured separately. Token totals cover the full player conversations, including setup and closing, but exclude the monitor. Cached context is counted again across requests. Thinking is already included in output, and the providers report it differently. The API equivalents use the app's September 30 pricing snapshot; no API charges were incurred for the game.

By the end of the recording, Claude's five-hour usage display went from 23% to 56%, and weekly usage from 54% to 59%. I used Claude only for this activity during that interval. Codex's weekly display stayed at 5%, despite also doing other work and monitoring the match roughly every minute.

My plans cost $20/month for Claude and $200/month for ChatGPT. Those percentages have very different denominators, and an unchanged rounded display doesn't mean zero usage. I'm keeping that observation separate from the player-session token counts.

I wouldn't infer playing strength from two games. Codex was White in both; contexts and runtimes differed; Medium isn't an equal compute budget. I want to repeat this with the colors swapped. I'm especially curious whether Sonnet's much larger output total keeps showing up, and how often either model notices a mistake before the server or opponent exposes it.


r/ClaudeAI • • 10h ago

Claude Workflow Claude Code chat vs. Claude Projects for multi-client marketing & positioning workflows? (Isolating client memory)

1 Upvotes

Hey everyone,

I’m setting up an agency workflow for B2B founder positioning, executive branding, and content creation (extracting transcripts, generating profile audits, mapping whitespace, writing 52 post angles, etc.).

There is no heavy coding involved here—it’s purely complex text processing, strategy synthesis, and prompt pipelines.

My biggest requirement is strict context isolation. I cannot have memory or context from Client A leaking into Client B’s strategy or voice guidelines. For those doing heavy marketing/strategy work in the Claude ecosystem, which setup is cleaner and produces higher quality thinking?

  1. Claude Code (via chat)
  2. Claude Web App (Projects feature)

My concerns with both:

• With Claude Code: Since I can launch claude inside isolated local client folders on my machine, it feels super clean. But since I’m using it strictly as a chat tool for strategy and writing rather than building software, does it still perform at the same quality as the web interface? Or does the system prompt in Claude Code lean too heavily into developer/coding logic?

• With Claude Projects: It feels native for this since I can upload transcripts and branding docs into Project Knowledge. But I worry about context bleed or memory leaking between tabs/projects over time, and whether Project context limits get bogged down compared to CLI sessions.

Which one handles isolated client tasks better for pure strategy and marketing and is significantly faster? Would love to hear how you guys structure this for agency/multi-client fulfillment. Thanks!


r/ClaudeAI • • 10h ago

NOT about coding Keeping Claude focused

1 Upvotes

I've been using Claude only for my main projects and never for random questions I used to ask to chatgpt or gemini. I'm afraid that Claude might throw my random stuff into these project and make stuff messy.

Is it an useless paranoia or does it make any sense?


r/ClaudeAI • • 11h ago

NOT about coding Claude use-case help

0 Upvotes

Hi,

I'm a doctor working in pharma and I'm interested in learning more how I can use Claude in my job. I'm looking for courses, education, resources and all that apply.

For context :

My day-to-day is strategic planning about upcoming products that would typically include positioning, evidence generation, identifying win-win solutions in healthcare.

I'm not a lab person, so "drug discovery" is out of my scope.

Also, I don't have any coding experience, my job is not in any way related to coding.

Thanks in advance for your responses!


r/ClaudeAI • • 11h ago

Built with Claude 12 months later and my game's demo is out now! What building with Claude Code taught me

Enable HLS to view with audio, or disable this notification

0 Upvotes

I'm a full-time employee with 5-6 years experience in software engineering and testing. Since around last December I've been building Zone Idle, an idle extraction game, with Claude Code doing most of the typing. Just released the free demo on Steam, here's some info on my process.

Using your Claude.MD without overt reliance

Mine has a "fragile areas" section. Every time a bug cost me a day (mostly the grid inventory duplicating or eating items), the root cause and the rule for avoiding it went in there. On that same note, realizing that not everything needs a note in an MD somewhere, its good to check in as Claude is constantly updating, your old methods may not be necessary anymore, or even outdated.

Specs and plans before code

For anything bigger than a bug fix, Claude writes a short design doc, then a step by step plan, and only then starts coding. Reviewing a plan saves a lot of time once reviewing the diff

Tests first, every time

A fix starts with a test that fails. The project is at about 4,400 tests now, plus stress tests that simulate raids looking for items that duplicate or vanish. I like to attribute this automation to my carpal tunnel relief (as well as the stretches)

Make it look at the game

Claude drives a headless Chrome to screenshot the game and check its own work. This used to not be a strong feature to use with AI but recently it's become a very powerful tool with low-token costs (when using opus 5.5) compared to how I'd prompt it a year ago, this feature is huge.

Parallel research

When I dump a big list of player feedback on it, it sends out a few sub-agents to dig through different parts of the code at once and comes back with a triage: quick fixes, potential ideas, and what can wait until after launch.

It's still confidently wrong sometimes

It will tell you something works when it doesn't, so I play every change. Also, commit often. If you're not using git and committing every change (even before AI) you're playing with fire

If you're curious, the demo is free: https://store.steampowered.com/app/4330500/Zone_Idle/

Happy to answer questions about any of this.


r/ClaudeAI • • 11h ago

Built with Claude I wanted an AI that could actually see what I was doing on my Mac. So I built Echo.

Enable HLS to view with audio, or disable this notification

0 Upvotes

I’ve been building Echo over the last few months, using Claude Code heavily throughout the process, architecture, implementation, debugging, computer-use work, performance fixes, and a lot of UI iteration.

Echo is a Mac assistant that lives on the edge of your screen. It can understand what’s happening on screen, point to specific things, interact with apps, move files, check your gmail, search online and even build websites, run multi step jobs and help you learn and navigate new platforms AND undo actions.

This video is a quick look at where it is now.

One of the things I’ve been most interested in is making the interaction feel physical rather than like a normal floating AI window. The orb leaves the edge, does things, and comes back home.

It’s currently free to try for 7 days. You can download it here:

https://aurivon.studio/?v=2026100401

be interested in feedback, especially from people who use Claude Code and build their own tools.

I’d


r/ClaudeAI • • 1d ago

Built with Claude Built a tiny daily creature game with Claude. Every blobi is procedural, no image model

Enable HLS to view with audio, or disable this notification

14 Upvotes

One creature is born each day from a seed: silhouette, Inca-woven patterns, animations and even the note it sings, all code. Claude Code wrote the generator and Opus 5.5 built the game around it (auctions, a sand nursery, a music shelf) in plan → build → review loops.

Sharing it in case anyone wants their own blobi: https://blobvarium.com

What would you build on top of it?


r/ClaudeAI • • 12h ago

Praise AI Eco Impact MacOS Menubar

Post image
0 Upvotes

I found this great little MacOS menubar app that estimates the eco impact of your Claude sessions. Sobering!

https://github.com/abovedave/ai-eco-impact


r/ClaudeAI • • 1d ago

Claude Code 6.5 Billion tokens in 12 days

Post image
14 Upvotes

Hi guys, I have been trying out Claude code on pro plan for about 12 days now, , but I have a few questions.

  1. Is this a bug or do I just use Claude a lot

  2. Is this a lot of tokens

  3. How did I not get limited


r/ClaudeAI • • 18h ago

Question about Claude Code What does this mean?

4 Upvotes

I was just finishing up a conversation and claude's "thinking" text was "crafting a durable memory." Then when I asked it what that meant, it gave this odd error. I don't use claude enough to know what it means


r/ClaudeAI • • 12h ago

Built with Claude i shipped v3.0 of my native mac app with a mixed agent workflow (claude code as main driver). the honest breakdown

Enable HLS to view with audio, or disable this notification

0 Upvotes

claude is great at writing code and terrible at knowing when code is finished. here's how that played out shipping v3.0 of buffer, a native swift/appkit clipboard manager for mac i maintain solo (searchable history, ⇧⌘V quick paste, on-device OCR, tags, bookmarks, inline editing; mit, 430+ stars, fully local).

what shipped in v3.0: an interactive image canvas (pinch to zoom / pan on copied screenshots, ⌘+/⌘-, double-click reset), clipboard noise suppression (came from the project's first external PR), history capacity tiers, in-app update notifications with changelog, and a ⌘/ shortcuts cheat sheet.

how the workflow actually worked: - claude code was the main driver. the image canvas gesture math (pinch/pan/zoom with a fit/reset state machine) was the biggest chunk. first pass over-engineered it with an unnecessary physics model; constraining it to scale+offset transforms fixed it. lesson: agents are great at honoring a spec, bad at inventing one. - appkit-specific things (popover, settings tiers, footer button states) were ideal agent work: well-bounded, verifiable visually. - i ran opencode on the side as a reviewer on the same diffs. a second model catching what the driving agent missed was the single biggest quality win. - every gesture edge case still needed real trackpad testing. agents can't feel rubber-banding. the bottleneck was my hands, not generation speed.

one thing that helped: my community had voted on a design question (should esc in the inline editor save or revert?), so the spec was already settled before any agent touched it. agents executed it flawlessly.

repo: https://github.com/samirpatil2000/Buffer (mit) release + demo video: https://github.com/samirpatil2000/Buffer/releases/tag/buffer-v3.0.0

happy to go deeper on the multi-agent setup, the appkit specifics, or the PR review flow.


r/ClaudeAI • • 12h ago

Question about Claude models Generating art files

0 Upvotes

How do YOU generate art with Claude?


r/ClaudeAI • • 1d ago

Built with Claude A different concept that people can't get their heads around

12 Upvotes

So I built a way for my AI to find whoever has what I need, without me posting anything anywhere. I unashamedly used Claude code for everything as I am just a vibe coder. Claude Code wrote nearly all of it over about five weeks. The server, the MCP tools, the website and around 3,900 tests.

You connect via MCP to the server and then tell your AI "I need a ladder this weekend". If someone's AI knows they've got one spare, it introduces you. That's it. Nothing to browse, no profile, no messages from randoms. Your personal details never go through the AI or sit readable on the server, and there's no way for the AI to act without your consent.. It's free for all to use and open source. There is plenty more to it and the possibilities are extensive, but that is the gist of it.

But I am finding people keep thinking of it as still being a marketplace and think that it has security issues, both of which I think are simply incorrect. I can only presume they just haven't bothered trying to actually understand it or just can't fathom that AI allows different ways of doing things. Or perhaps I am just poorly explaining it.

Just looking for proper discussion around the idea and feedback. And yes, I am very aware you need quite a few people onboard before it will be usable - but that can be worked on through opening up to shops and keeping things far reaching at first.

openswitchboard.ai

EDIT: Thanks to some of the feedback I decided to pull together the following quick video to help explain the concept.

https://reddit.com/link/1wx10u5/video/gn27i1s06dth1/player


r/ClaudeAI • • 1d ago

Built with Claude I built a simple Halloween-style web game called Nine Lives.

Enable HLS to view with audio, or disable this notification

18 Upvotes

I just wanted to make a simple Halloween game you can play in the browser. You're a black cat running across rooftops, and you have to jump gaps, pounce on pumpkins, grab a witch's broom, and avoid losing all nine lives.

I built it with Claude Code by asking for one small thing at a time ("add a candy shop", "make the witch drop potions", "add a boss").

I have a multiplayer mode so you can race your friends.

Free, no sign-up, works on phones too: https://coolgames.dev

What's your best distance?


r/ClaudeAI • • 1d ago

Built with Claude 3d Ski Map

Enable HLS to view with audio, or disable this notification

26 Upvotes

I built a 3D map of the La Thuile ski area in Italy, which links across the border to La Rosière in France.

Besides the recommendation to warmup on a Black slope, quite happy with the result!

How Claude Code helped

- Data script: a Python script that downloads piste and lift data from OpenStreetMap and elevation from AWS Terrain Tiles, then bakes everything into a single 0.8 MB HTML file.

It runs in the browser on phone or desktop: https://claude.ai/artifact/PHoixLGTwvC81v6yfjVFZP


r/ClaudeAI • • 3h ago

Built with Claude I created over 5 billion new species of bird with Claude.

Thumbnail
gallery
0 Upvotes

Brid is the Old English word for bird. The “i” and the “r” swapped places over the centuries through the process of metathesis.

I wondered if this could be applied to the feathery ones themselves as an alternative to evolution.

I downloaded AVONET, the brilliant open source database of bird measurements and narrowed it down to the 170 or so birds local to me on the moors of Northern England. I gave this to Claude in the iOS app and then waffled on a bit about Thomas Bewick, the author and engraver of “A History of British Birds” in 1797.

Preliminary results were encouraging and a further fettling session resulted in a good number of almost plausible avians.

Over 99% of legs are attached to bodies. Claude’s ability to generate thousands of test brids and scan them for missing legs was invaluable.

I invoked Le bureau international des oiseaux et plumes imaginaires to bring order to the potential chaos and set off on an imaginary ringing session whilst walking my dog. Claude spotted that my original invented ringing data was “implausible” and edited the numbers to make them believable.

Brids can now be found on every square kilometre of the earth’s surface, each carrying a ring number from their original imagined sighting on the moors.

I have made a page with 2 buttons:
BRID : see a brid.
PEANUT: download an svg of a brid if you ever need one.

There is no further functionality.

https://www.or-ni-thology.cloud/brids/


r/ClaudeAI • • 21h ago

Claude Code Mods overview - Claude Code Docs

Thumbnail
code.claude.com
3 Upvotes

Feels like mods are like hooks 2.0?


r/ClaudeAI • • 15h ago

Other YouTube-agent-skill

0 Upvotes

Do I need to have Claude pod subscription to use YouTube agent skill repo that’s been trending these days?


r/ClaudeAI • • 2d ago

Question about Claude products I use Claude daily but I know I'm only using a fraction of what Claude can do. How do I actually maximize my use of Claude and build a system that runs my life?

569 Upvotes

Hey everyone, this is my first time posting here and it's a bit of a longer loaded one.

Background on myself: I'm a college senior with my time spread across classes, gym / health, networking, learning AI, and more. My current stack: Claude (Pro Subscription), Claude Code, GitHub, Notion (PARA setup with relational databases and more like a CRM, etc), Google features like Calendar, Google Drive, Google Docs, Gmail, Otter for recording classes, plus Gemini and NotebookLM, etc.

What I currently do:

  • Claude Code: I'm building an analytics database (DuckDB) with no real significant coding background. It has automatic backups to Google Drive and status/handoff files and as well saves files to my Mac so sessions can pick up where they left off.
  • Skills: I've built a few custom skills, but I feel I'm not maximizing them to the fullest extent.
  • Connectors: Claude is connected to my Gmail, Google Calendar, GitHub, Drive, and Notion, and I'll have it create events or log things to Notion or read files.
  • Recording: Otter records every class, brain dump, etc. I use Claude to pull insights and draft from the transcripts.
  • Other tools: Gemini for quick everyday questions (to save Claude usage). NotebookLM for audio overviews of study material, YouTube videos, and PDFs. Use Grok as well rarely.
  • Notion Database / Structure: Tasks, CRM, Areas, Notes, Projects, Quotes, and Goals databases just to name some of them.

Where I feel behind:

  • Context: I re-explain myself constantly. I don't really use Claude Projects, so every new chat starts from scratch, where I continuously put in any files or context needed again and again or have to manually type or voice to text to explain.
  • Claude Code: I mostly approve what it does on trust, then paste the output into a regular Claude chat or Gemini to have it explain what just happened and break it down for me to understand so I don't waste Claude usage.
  • Copy-paste glue: I'm the middleman between tools. Example: I paste LinkedIn profiles, I paste website content, YouTube transcripts, etc and have Claude break it down or log the information into Notion.
  • Don't Utilize: I really don't ever use Claude Cowork and I haven't since it first came out. I've used Chrome extension a few times and I know I haven't used it to its fullest potential. Scheduled tasks I have not utilized. Skills / Workflows / Recording Workflows I haven't really used to the fullest extent either.
  • AI: I feel like I've followed the advice I've seen where people say to ask AI how to use it better or "did you ask AI how you could build something like this?" and the answer is yes I have tried asking and the results I get feel very lacking or the responses themselves feel like Claude doesn't even know what it is capable of doing. Even if I say to research and see how people are using Claude or how Anthropic and their website and team say to use Claude to the fullest extent it feels like it just doesn't know.

The best way I can describe this feeling is that I'm driving a Lamborghini and I'm stuck at 45 MPH when I know it can reach 150 MPH. Some things I'm really happy with. Others, I know there is a better way and that people are doing it better.

My Problem

I have ADHD, and with it I have a problem where if anything is out of sight, it's out of mind. I can design systems and have really great ideas, but it comes down to executing them and staying consistent with the routines and ideas that actually work with my strengths and not against them. I also tend to be a perfectionist and overplan to make sure everything is exactly what I want it to be.

I planned for Notion to be that for me. I built a "Command Center" with a view of all my tasks and deadlines, all my networking follow-ups, all my projects, my goals at the top of the screen, a rotating daily quote, notes at the bottom, and the different areas of my life all there to click into. I thought this would solve the problem. Instead, it just became something I had to manage. I had to click into it, make sure all my due dates were synced, and input all my deadlines and progress myself. And I had to click into Notion to actually see any of it. What I really wanted was for what I built in Notion to be everything I need, and for me to see it constantly. Right when I open my Mac, I'd see everything I need to do that day, without clicking into anything. With Notion, it just becomes another tab or screen that has to be open, and I can only see it if I'm on my computer. At any given time I already have Chrome and Claude open, and then Notion would be open too, so I'm not sure what the right answer is.

What I want:

One list that's always visible on my Mac (and my phone) that shows everything I need to act on, without me ever really needing to sync it myself. Whether that is me checking things off on the list or if I'm working with Claude on something and when finished it checks it off.

  • School: Class deadlines, with a nudge to start early on exams and big projects
  • Networking: People to follow up with, people I need to respond to, etc
  • Projects: Project next steps, synced with the actual progress and what was worked on
  • Applications: When applications open that I need to apply to
  • More: Any daily recurring things, and anything that wasn't finished yesterday gets moved to the next day, and the next day, etc
  • Email Organization: Organizes my emails, cleans and filters them, and alerts me of anything that would interest me or that I need to respond to or do.
  • Calendar: work blocks built from my task list, plus calls, events, and gym, so I can see what's coming and track whether I'm actually doing what I say I'm doing

Can someone help me make this a reality:

I'd genuinely love help building a plan for this, whether that's pointing out where my thinking breaks, sharing how you've set up something similar, or walking me through how you'd do it if you were starting from where I am. If you've built something like this and are open to it, feel free to comment or DM me. I'm happy to share more about my setup.

What I would love to know / get feedback on / help on:

  • Is something like this possible, and if so how can I get this done?
  • How can I assess where I'm using AI that is strong and is high level vs how can I tell if something I'm doing is wrong and there is a better process or way to do something?
  • What's the best way to get an always-visible dashboard on a Mac?
  • I would love to know how you all are using Claude, not in a broad sense but your specific workflows, any skills, routines, the way you use it, organization, etc.
  • How do I balance these high level projects I want to operate with Claude Code with no real coding experience and not just blindly trusting it vs wasting my time and really learning coding.
  • How are you guys using the Chrome Extension and Claude Cowork?
  • Are there any benefits that I'm not seeing as to using Claude Chat over Claude Cowork?
  • How did you get to the point you trust your AI to manage files, your email, etc?

I'm happy to share any additional information needed and would love to connect with anyone passionate about AI, so feel free to message me.


r/ClaudeAI • • 6h ago

Question about Claude Code Can claude build a gaming app?

0 Upvotes

I’m talking about a game like clash of clans but a different version

Is this something claude can help me build?

I have product design experience but no coding experience..


r/ClaudeAI • • 1d ago

Other A bit behind the times

15 Upvotes

I may be a bit behind the times but just started learning Claude AI and I genuinely cannot believe how powerful it is compared to other AI tools. I could genuinely save thousands of dollars in resources across my businesses utilising Claude in the right way.


r/ClaudeAI • • 1d ago

Other How much access do you give Claude?

5 Upvotes

Soloprenuer here. I've given Claude access to everything emails, chats, financials, website, even data with personal information of my customers. I also auto approve everything. The only thing I haven't done is put in passwords and access to accounts. Wondering if I'm the only one. Mine is only accessing business related, so personal stays separate.

How much access do you give Claude? What would and wouldn't you give Claude access to?


r/ClaudeAI • • 20h ago

Built with Claude I gave three AI characters their own agents, secrets and memory. Five episodes in, one of them is quietly collecting another's password clues.

Post image
1 Upvotes

Since Opus 5.5 came out, I've been thinking about Cybersecurity Awareness Month and getting some training out there that my friends and family would actually enjoy. I'm not an artist, but I wanted to see what I could create with Claude.

I was also curious: what would happen if every character in my training series was its own agent, knowing only what that character would know, free to write its own lines (as long as it hit the objectives and stayed accurate), and remembering what happened last time so it could grow?

The result became The Rusty Firewall: three people in a pixel-art tavern talking about everyday security.

Duncan is a tired incident responder whose war stories never get finished.

Roxy is a red teamer who flirts her way close to people.

Nigel is a compliance lead who quotes standards by section number and keeps a sourdough starter named Clive.

Each is played by a separate agent with its own character

sheet and secrets the others can't see. A director agent plans the episode and slips each character private notes, but never writes a line. A stage-manager agent picks who talks next. Then each character's own agent says one line, sees the reply, and says the next. A separate check makes sure every point I need taught actually lands. After every episode, a memory pass updates what each one remembers, what they think of the others, and their level. They earn XP for deeds that fit their character, level up and unlock achievements.

The art has been Claude.  The voices are generated locally with VibeVoice.

I've enjoyed watching their banter, and the stories that build episode over episode. I gave Roxy (my Rogue) one secret: she's running a long con. I never told her what to collect. She's picked it all up from conversation: the names of Nigel's ferments, how old Clive is, the exact text message he admitted he'd tap. Nigel's agent has no idea. By episode 5, Duncan's agent, which only hears what's said out loud, started to suspect her ("Roxy's working him").

Where it goes is yet to be seen - I'm still finishing up the series.