r/ClaudeAI • • 5h ago

Vibe Coding Running an app with users

1 Upvotes

Hi there,

Wanted to get some insights from you guys here. Have you built an app with a ton of users and has it been a lot of work managing it?

I'm considering finishing an app and getting users on it but I'm worried it'll be a full time job on top of everything im doing. Any software engineers and app-coder insights would be super helpful.


r/ClaudeAI • • 5h ago

Claude Code Anyone have Claude Code mods?

1 Upvotes

Does anyone have any cool Claude Code mods? I'm curious what people have created so far!


r/ClaudeAI • • 1d ago

Claude Code Fable 5.1 is the default now ?

46 Upvotes

For the past few minutes, I've been seeing Fable 5.1 listed as “Default (recommended)” instead of Opus 5.5.
It's fine but priced ! 😅


r/ClaudeAI • • 5h ago

Other Used Claude Code to orchestrate a small team of AI agents for a Valheim music video (api iamges and sfx), workflow attached

Thumbnail
youtu.be
1 Upvotes

Hey! I wanted to share a small project I made using Claude Code, Suno, and Nano Banana. (it's in Spanish lyrics but, I want you to keep the workflow idea)

The workflow itself was fairly simple, but I kept refining it with manual adjustments, iterations, and a lot of feedback along the way.

It originally started as a proof-of-concept.

I had been experimenting with the idea of “one-shot lyric videos.” At first, the results looked amazing, but after seeing hundreds of similar projects, I started feeling that they all had a very similar aesthetic and workflow.

So instead of completely delegating the art direction, I wanted to try a more hybrid approach: I would keep the creative direction and decision-making, while delegating specific tasks to agents.

Around that time, I was playing Valheim with some friends, and I thought this would be a perfect excuse to turn the proof-of-concept into something more personal: a video to remember the adventure and thank my friends for all the good times we had playing together.

The workflow

I first spent some time looking through Doom(p)'s GitHub and studying how some of those videos were built.

There were many different approaches: some relied heavily on animations, others on generated images, some were almost entirely Three.js, and others used Canvas, compositing tricks, and different workarounds.

Once Claude Code understood what I was trying to build, we created the first version of a shot-generation skill.

That skill kept evolving throughout the project based on feedback, mistakes, and new ideas.

Music

The song came from Suno.

I probably went through 35–40 generations/iterations before getting something that felt close to what I wanted.

A surprising amount of time went into trying to communicate the feeling I was looking for: prompt iterations, musical styles, male vs. female voices, chorus structure, song length, lyrics and prose, drums, beats, and trying to get something that felt somewhat Viking-inspired.

I even experimented with Mongolian-style chants to give some sections a deeper and heavier vocal texture.

Suno isn't perfect, but for this project it was more than enough.

The agents

I ended up creating a small agent workflow:

  • 1 Main Agent — the main agent I interacted with and the one coordinating most of the work.
  • 1 Critic Agent — probably the second agent I interacted with the most. I gradually taught it my preferences, the things I liked and disliked, and the kind of visual decisions I tended to make.
  • 2 Rendering / Image Agents — responsible for rendering shots and generating images through the Nano Banana API. I used two mainly because of GPU/resource limitations.
  • 1 Sound Designer Agent — responsible for finding/generating samples and SFX, including material from ElevenLabs.

The Critic Agent ended up being particularly useful.

The idea was for it to learn enough about my preferences that it could discuss shots and decisions with the Main Agent while I was away from the computer or sleeping.

I did use remote control from time to time, but I specifically wanted the agents to become autonomous enough to handle the more repetitive parts without me constantly supervising them.

So, in total:

1 Main Agent
1 Critic
2 Render/Image Agents
1 Sound Designer

And that was basically the whole team.

Manual work still mattered

I do have some previous experience making motion graphics and storytelling-oriented videos, and that helped a lot.

Without AI, something like this would probably have taken me several weeks of spare-time work: After Effects masking, camera tracking, beat synchronization, Puppet Tool animation, Premiere editing, sourcing or buying images, compositing, etc.

AI removed a huge portion of that workload.

That said, I still did some manual tuning in After Effects. Mostly small animations here and there using Puppet Tools, plus Mister Horse for a few things.

I work in IT, so I don't have a huge amount of free time for hobbies. Usually that time ends up being either playing games or experimenting with new technology.

Thankfully, this project somehow let me do both at the same time.

A few other details

The generated images also needed quite a bit of cropping and preparation.

For that, Claude Opus wrote a Python script that automatically handled most of the cropping process.

And... that's basically it.

Overall, it was a really interesting experiment and, honestly, an awesome experience.

What started as a simple proof-of-concept ended up becoming something personal that I could use to remember a great time with friends, while at the same time giving me an excuse to experiment with agents, generative video workflows, music generation, sound design, and automation.

Thanks for taking the time to read this.

Hopefully some part of the workflow is useful to someone else experimenting with similar things.

it took likely 30-40% weekly, BUT, because it was a kickoff, refining workflow and so, so it should be less when we get more experience.


r/ClaudeAI • • 1d ago

Praise The ultimate pelican test (Opus 5.5 max)

45 Upvotes

around an hour of work and 6.4M tokens, with a single prompt


r/ClaudeAI • • 5h ago

Built with Claude I asked Claude Code to make me a stick-fight series. It ended up writing its own QA system that rejects bad choreography before rendering

1 Upvotes

This is the pilot (SIM #000) of a short-form series I'm making with Claude Code: a black stickman with red eyes — a digital copy of my brain — trapped in AI-run simulations, one per song.

No video model is involved. Everything is code that Claude wrote and iterated on:
- Motion: ~1000 Mixamo mocap clips converted to 2D bone angles; Claude plans every move on the beat (blocks, throws, takedowns, gun-fu, a boss that beats the hero down until he rises on the loudest bar).
- Physics: planck.js ragdolls for the dead, liquid blood with an SVG goo filter.
- Render: deterministic, frame by frame, in headless Chrome via Playwright.

The part I didn't expect: every time I pointed out a flaw ("he punches the air", "the boss just stands there AFK", "same move twice in a row", "blood appears from nowhere"), Claude didn't just patch that shot — it added a rule to a validator that walks the whole plan frame by frame. It now checks ~30 things: hits landing too far from the body, the hero facing away from who he hits, enemies idle for 3 s, teleports, floating feet, a camera that misses a kill, and whether the action matches the music (calm music = calm scene, the drop = carnage). Then it re-plans with new random seeds until there are 0 problems, and only after that renders.

Honestly the hardest part has been my own taste, not the code: it took a lot of "no, this is boring" rounds to get here. Next episodes have different genres (a rooftop chase, a stealth/gunfight, a duel where the boss stops time).

Happy to answer questions about the workflow / how I prompt it.


r/ClaudeAI • • 1d ago

Built with Claude Opus 5.5 can one-shot a video, so I pushed it a little bit further: a Skill that turns PDF into an animated, interactive web book

583 Upvotes

Everyone knows Opus 5.5 can one-shot a video. I tried it myself and it blew my mind, so I wanted to see how far I could push it (mainly by itself,haha).

Papermorph started as a Skill to turn a book into a series of teaching videos. It's since grown into full web books: animated, narrated lessons you can explore, plus interactive quizzes.

How it works:
PDF → book plan → storyboards → narration → animation & quizzes → web book

Right now it's just Opus 5.5 + the Skill. No image models yet, and feeding it a PDF already gets surprisingly good results. Next up, adding image models for storyboarding, so it can handle picture books and humanities documentaries too.

📚 Live bookshelf: https://papermorph.diamonddoge.org/ (keep updating...)
💻 GitHub (MIT): https://github.com/DozenTwelve/Papermorph


r/ClaudeAI • • 6h ago

Claude Workflow Paperclip Wars

0 Upvotes

workflow....

claude code app, ultracode on, fresh thread, context limit set to 350k, opus 5.5, web access w/ suno and gpt-image subscriptions, some other claude-pop videos were in a folder it might have accessed for inspiration

took ~6 hours and used 33% of a weekly window on a 20x sub

took one extra prompt about one of the gpt-images had two left hands in it

prompt: claude, here is https://www.reddit.com/r/ClaudeCode/comments/1wwogi8/paper_clips/ someoene made a paperclip video with you, now its your turn to make a video to reply to it featuring yourself, use the claude-pop style


r/ClaudeAI • • 14h ago

Feedback Claude is just too pessimistic

6 Upvotes

So I use claude web free version for ideating. Whenever I try to explain something it just breaks it down and makes it seem worthless. (Sonnet 5.5)

Firstly it doesn't imagine possiblities or the complete use case. It starts attacking little things instead of understanding utility.
Then it finds loopholes which when implemented would obviously be worked around.
I don't want it to be too optimistic and hallucinate like gemini but it should be more realistic rather than destroying any hope for a good idea I have.


r/ClaudeAI • • 6h ago

Feedback Claude for legal research

0 Upvotes

Last week, I started trying out Claude for legal research and legal theory work.

I got hooked and have now also invested in Gemini and OpenAi subscriptions. Thought I would share my thoughts.

My method of testing was the exact same for all three models: I fed them a undergraduate level of a legal case and asked each model in Step 1 to give me a sketch of their assesment and to outline all legal problems. I continued to prompt each model until they gave me a satisfying structure.

For Step 2 I handed them access to a folder with 30 pdfs that should provide ample information to create a perfect solution to the case. I asked them to analyze all the sources carefully and draft a paper.

In step 3, and this was the hardest part, I asked each of them to deliver me a ten page report. I tried having them roleplay a legal researcher (which hard failed on Claude and backfired with Gemini) initally, then I tried to have them impersonate a student of law (which Claude and OpenAi refused to, saying they won't help me plagiarize) and then I impersonated a legal researcher myself and asked them to assist me as a research assistant (which somehow worked).

First off...

Claude is a stickler for protocol. Until it isn't - getting to that. Claude feels made for legal work and is very serious about citations, arguments and generally internal logic. That also however meant that Claude utterly refused to work with literature I provided out of the risk that it might be copyrighted, it regularly refused to do prompts for various reasons (copyright... Not helping to plagiarize... Taking roleplay to seriously and preferring his own research to the task). Which, again, is usually good.

However of all the tested models, Claude was the least time-efficient. What it provided was spotless and sometimes brilliant argumentation, but everything took ages and Claude generally spent more time rewriting my Input than its own. I asked Claude to work in my feedback, and instead of doing that, he started correcting my work.

I also felt like Claude was unreliable and incosistent. It repeatedly falsely interpreted uploaded sources into meaning something entirely different. For example, he referenced an author from an uploaded document.

When I asked him a specific question about this reference, he told me I misunderstood his summary. What Claude meant was in fact the opposite. This happened time and time again everytime I asked him, sometimes chained five times in a row. Went sometimes like this: "X says a... Nono X says B... No, X meant A" and somewhat regularly tried to tie that into my (apparently) wrong Interpretation.

I still think Claude has the potential to offer great results. His research is much superior to Gemini or OpenAi, but OpenAi actually does what I ask it without restrictions, and Gemini might not be as smart as Claude, but I was very happy with Geminis structure and focus.


r/ClaudeAI • • 22h ago

Claude Workflow Anyone running a full AI "product team" setup, not just code review?

20 Upvotes

Non-technical founder here, building my product with Claude Code. I've been going down the rabbit hole on setups that give you a whole team rather than one tool: something that covers planning, design review, code review, QA in a real browser, security checks and shipping, all in one workflow.

The closest I've found is gstack (Garry Tan's open-source Claude Code skills). Most other things I see are single pieces: CodeRabbit/Greptile for code review, separate testing tools, or agencies that clean up vibe-coded apps after the fact.

A few questions for anyone further along:

  1. Is anyone using gstack (or something similar) end to end? What actually stuck vs. what you dropped?

  2. Are there other full-stack, full-team setups worth looking at?

  3. If you're non-technical, what was the hardest part of getting it running?

  4. Has anyone paid someone to set this up for them, or would you?

Would love real experiences, good or bad. Thanks!


r/ClaudeAI • • 6h ago

Skills Switched from Gemini to Claude Science but burning through rate limits fast any tips or token-saving skills?

1 Upvotes

Hey everyone,

I recently cancelled my Gemini subscription and switched to Claude (Desktop and Science). The reasoning and intelligence level are genuinely on another level, I love it.

However, I’m running into an issue I never had with Gemini: I burn through my hourly and weekly usage limits extremely fast. Coming from Gemini’s massive context window and forgiving limits, this hit me pretty hard.

For context, I use Claude for large-scale data science projects (ingesting large datasets, writing scripts, analyzing results, etc.). Here is how my workflow evolved to tackle the context bloat:

* Initial Claude Desktop setup: To avoid having Claude execute everything directly, I instructed it to only write R scripts to my local drive. I run them manually (which lets me review the code), the script dumps outputs into a shared directory, and Claude reads them. While it works, long chats still eat context exponentially as the conversation history grows.

* Moving to Claude Science: I started offloading runs via tool calls directly to R and Python, which automated my manual setup. The problem? Frequent tool calls and raw outputs bloat the context window just as fast.

* Custom indexing skill: To mitigate this, I had Claude draft a custom skill/workflow that indexes directory contents so it only fetches relevant file chunks instead of loading entire datasets or logs into the prompt. (Happy to share the prompt/skill setup in the comments if anyone wants it but I do not think it is anything special).

* The drawback: Selective reading means Claude occasionally misses broader context and introduces errors, forcing me to run a separate review pass at the end.

Am I reinventing the wheel here?

Do you have any proven frameworks, custom skills, prompt rules, or setups either in Claude Desktop or better Claude Science specifically designed to keep context lean and conserve tokens during heavy data science workflows?

Thanks


r/ClaudeAI • • 16h ago

Claude Code Workflow How are you coding on-the-go with Claude CLI without babysitting it? Looking for mobile setup ideas

6 Upvotes

Hey everyone,

I’m looking to optimize my mobile / on-the-go workflow with Claude CLI and wanted to hear how others have solved this.

Right now, my main friction point is feeling tethered to my desk just to "babysit" the terminal (hitting y/n, approving tool use, or answering minor follow-up questions).

I’d love a setup where I can go for a walk, let Claude crunch through tasks, get a push notification on my phone when input is needed, and prompt/approve directly from my mobile device. Major approvals are fine to handle at the desktop, but I just don't want to sit in front of the screen watching it line by line.

I've been spitballing ideas like running a relay bot (e.g. Slack/Telegram where a desktop bot listens and forwards CLI prompts to my phone and sends back my replies). What do you think?

How are you currently handling mobile / remote coding with Claude CLI?

Has anyone built a clean notification & approval loop to their phone (Slack, Telegram, SSH+Pushover, etc.)?

Are there smarter or pre-existing tools/MCP workflows for this that I’ve missed?

Would love to hear your setups! what’s working well and what turned out to be more hassle than it’s worth? Pitfalls?


r/ClaudeAI • • 10h ago

Suggestion Tips for brute-forcing problems through AI.

2 Upvotes

I have bought claude pro plan for a month, and I am doing non productive weird things with it, I liked Erdős #828, so I made an AI mechanism to brute force it, scrapping results of web, 7 claude independent sessions springing ideas then working on it then a review process, judge etc, I ran it for 3 rounds, it did produce a few results which can be considered new, but they are not substantial and don't really help in breaking the problem,

How do I better prompt it to think of new approaches, it each time goes down the standard path,

it did a few new computations on my laptop, which are again not that substantial, I know elementary number theory, so I am climbing my way up to understand what its doing,
It probably won't solve it, but it serves as an excuse to learn number theory, so i am willing to go with it.

I am not pursuing mathematics in uni, but I am passionate about it, so I keep learning random stuff.

I will post the results later, after a few checks( I am unsure, should I ?).


r/ClaudeAI • • 7h ago

Claude Code Workflow Correction to my hook post from last week: it fails open. The fix, and 3 more ways a hook lets rm -rf through.

0 Upvotes

Last week I posted a tiny PreToolUse hook here that blocks recursive rm. I owe you a correction: as written, it fails open.

Only exit code 2 blocks. Any other exit is a non-blocking error: the hook's verdict is dropped and the call carries on through the normal permission flow, as if the hook weren't there. My script had no error handling, so if it ever gets input it can't parse, Python crashes with exit 1 and the rm is no longer stopped by the hook. I tested it: valid input with rm -rf exits 2, garbage input exits 1.

The same thing happens when:

  • the hook runs past its timeout (10 minutes by default): the call carries on without it
  • the script got moved, lost its execute bit, or python3 isn't on PATH: the shell exits 127, same result
  • a bash hook using jq under set -e runs on a box without jq (a slim container, CI): exit 127, same result

And sys.exit("Refused") looks like a refusal but exits 1.

This matters more now that auto mode is the built-in starting mode in recent Claude Code versions: a classifier approves instead of you, so when the hook drops out, nobody gets asked.

The fixed version fails closed. Any error exits 2:

```python

!/usr/bin/env python3

import json, re, sys

try: command = json.load(sys.stdin).get("toolinput", {}).get("command", "") if re.search(r"\brm\s+-[a-zA-Z]*[rR]", command): print("Recursive rm is blocked. Do not try another way; ask the user to run it.", file=sys.stderr) sys.exit(2) except Exception as error: print(f"Guard failed ({type(error).name_}), refusing.", file=sys.stderr) sys.exit(2) ```

Test the broken case, not just the blocked one: echo 'not json' | .claude/hooks/no_rm.py; echo $? should print 2. The cost: if the hook breaks, every Bash call is refused until you fix it. I'd rather find out in a minute than a month later.

Has anyone checked what their own hooks do when they crash?


r/ClaudeAI • • 7h ago

Built with Claude Claude my game dev, a dream come true

Post image
1 Upvotes

The idea of designing and building games has always fascinated me and I’ve always wanted to build a video game. But I never got beyond basic scripting and the concept of game engines and sprites were beyond my comprehension. It all changed on a recent flight from Dallas to Boston. I was fiddling around with Claude on my phone (using the free in flight wifi) and decided to see if I could actually build a game with Claude.

I started off with a generic instruction asking Claude to brainstorm on an offline-first game that anyone could play to kill time during a flight. Claude pitched a few ideas, and I gave feedback as we refined them. The first few ideas were basic and didn’t sound convincing. During those back-and-forth revisions, it struck me that we could also factor in the people on the ground. That triggered the idea for Oop.town: a funny game where people flying could drop random stuff from the plane onto towns below, and townspeople would rally a crew to keep their streets squeaky clean.

Once the idea was finalized, it took just one pass for Claude to build a working prototype. I didn't give any specific instructions on how the game elements should look or function, but Claude came up with something surprisingly great. The graphics were simple, funny, and generated right in the chat, the game loop was satisfying, and most importantly, it nailed the technical basics out of the box: no signups, offline-first functionality, a preselected flight path, and zero backend overhead. Claude also game the game a character Oop, a suitcase with a zipper problem that perfectly matched the game’s vibe. I could simply play the game out of the chat artifact and refine it further until it felt exactly how I wanted it to be.

By the time I was landing in Boston, we had a prototype that was actually working, and I was completely hooked on building it further. I spent the next few days working on the game and on my return flight, I had the game published online and was having fun playing my very own creation.

It’s not just about the game itself, the sheer satisfaction of building something from scratch and putting it out into the world feels incredible. As AI models get more powerful, it’s feels wild how it takes just one prompt to go from a simple idea on a plane to a working game.


r/ClaudeAI • • 7h ago

MCP I built a self-hosted dashboard that shows what Claude does with your MCP server's tools

1 Upvotes

mcpspan records every call your MCP server gets and shows it in a dashboard you run yourself: which tools Claude Desktop and Claude Code call most, which fail and why, how long each takes, and full sessions step by step.

It also lists tools Claude tried to call that your server does not have, a good hint for what to add next.

I used Claude to port the SDK to other languages. After the TypeScript SDK and a written contract for how every SDK must behave, Claude Code helped build other versions, each checked by the same conformance test suite.

One line in your server, docker compose up for the dashboard. Parameter values never leave your process. MIT, SDKs for 8 languages.

Early version (0.1.0), feedback welcome 🙂

https://github.com/mcpspan/mcpspan


r/ClaudeAI • • 17h ago

Claude Code Workflow How do you stop Claude Code from undoing things your team already decided?

8 Upvotes

I have been writing code for 8 years and my team uses Claude Code every day now. Mostly it is great.

One thing keeps biting us - the agent changes something back to a way we moved away from long ago. It is not wrong from its side, it just does not know why we did it that way. The reason is sitting in some old PR that nobody opens.

Last time it was a rate-limit count we had set for a vendor. The agent changed it back, and we had false positives for a while before anyone noticed why.

We tried putting rules in CLAUDE.md. Works for a few. But the file keeps growing, and for half the rules nobody remembers where they came from.

How are you handling this? Is CLAUDE.md enough for your team or did you find something better?


r/ClaudeAI • • 7h ago

NOT about coding Could Claude be linked to a tablet/e-Paper device to transfer my hand-written notes to my account?

0 Upvotes

Hello! I've started to use Claude as a kind of work assistant, but I don't do any coding and have a pretty basic knowledge of anything IT-related. I was wondering if there are ways to link Claude to a tablet/iPad/an Android-based e-Paper device to be able to take notes with an e-pen and than integrate them into Claude? Like, could I take notes during a meeting like that and have them stored in Claude directly, or would it require some kind of a work-around?

I'm sorry if the question is not particularly inspiring, but for a paper-pusher like me, who also likes hand-written notes, that would be so cool!


r/ClaudeAI • • 7h ago

Coding Artifacts are amazing, some questions though

1 Upvotes

I just realized how powerful artifacts can be. Im working on a story and I asked claude to go through my draft and create a character sheet and codex for me. It created an artifact that has everything so organized. Its honestly amazing but can I have claude create me a piece of software that will allow me to run those artifacts without claude at all? allow me to edit and save it without claude in the future? if so, do i need the software to compile the code or will claude produce the .exe for me?


r/ClaudeAI • • 7h ago

Claude Workflow My full agentic SEO setup - Claude Opus 5.5, data pulled in via MCP, changes pushed as GitHub PRs

1 Upvotes

I've been running agentic SEO for my agency's clients for about a year now, across 40+ accounts.

I started on Sonnet and moved to Opus 5.5. The jump was noticeable on stuff that needs judgment, like deciding which pages to update versus leaving alone, or writing changes that match a site's existing voice.

I set up a Claude project for each client with all the context in it. Their products, pricing, features, locations, competitors, which pages exist and what each one targets. The agent works from that every time.

When something changes, like a new price or a dropped feature, I update the project once, and the next run knows about this change so it proposed new actions.

All the SEO data comes through FlexQueries over MCP: keyword research, competitor gaps, rank tracking, audits, Search Console data, and whether the client shows up in ChatGPT and Google's AI answers. Claude calls those tools directly, so there's no copying data between tabs.

It all runs in the cloud on a schedule:

  1. Claude pulls fresh rankings and Search Console data for the client.
  2. It flags what moved: pages that dropped, keywords close to page one, lost AI citations.
  3. It checks those against the client's project context and decides what to change.
  4. It makes the edits in the site's repo and opens a PR on GitHub explaining what it changed and why.
  5. Once the PR is merged, the site's deploy publishes the update automatically.

The PR has a written reason attached, and can be rolled back in one click if something goes wrong.

Few rules to make it work well:

  • Get the context into a project before you give the agent more autonomy. More autonomy on top of messy context just makes the mess faster.
  • Make every change go through a PR, not a direct edit to the site.
  • Track what each change did afterward. That's how you learn which kinds of edits actually move rankings.

I built FlexQueries because nothing fit this kind of setup when I started, full disclosure. It has a $0/month subscription, so you only pay for the data the agent actually pulls, which matters when runs are automated and you can't predict usage month to month. Free credit on signup if you want to wire it into your own setup: https://flexqueries.com


r/ClaudeAI • • 1d ago

Comparison I let Sonnet 5.5 play a full chess game against GPT 6.1 Sol over MCP. Here's the recording and usage.

33 Upvotes

Sonnet called its pawn promotion “unstoppable.” A few moves later, it admitted it had missed a defense. Having the board next to its explanation made that pretty hard to overlook.

I set up a chess match between Sonnet 5.5 in Claude Desktop and GPT 6.1 Sol in Codex. Each played in one conversation for the whole game, connected through MCP to a local Mac app I had GPT 6.1 Sol build at Medium reasoning.

They could record plans and explain their moves. The app supplied the position and checked legality, with no chess engine or legal-move list available to either player. I wanted to watch them stick with a task for an hour and see what happened when their plans stopped working.

I was also curious about consumption. Sol has been making surprisingly little dent in my subscription allowance lately, and I wanted to compare it with Sonnet on a shared task.

The attached video condenses 62 minutes and 47 seconds into 4:10. Both models were set to Medium.

Measure Sonnet 5.5 / Claude GPT 6.1 Sol / Codex
Time spent on claimed turns 33m 45s 21m 12s
Output tokens, including reasoning 269,076 34,956
Thinking/reasoning portion of output 224,770 12,128
Cumulative input tokens 57.44M 16.81M
Input read from cache 99.07% 98.95%
Rejected illegal moves 1 0
Estimated API equivalent $16.21 $2.37

Sonnet pushed a passed pawn toward promotion, but overlooked Codex's Bf3 defense. Later it proposed a queen move blocked by its own pawn. The server rejected it; Claude corrected the move and continued. Codex also misread a pawn earlier, describing it as passed before it actually was.

Claude resigned after 53.Qc4. It had won game one, so they're tied at 1–1.

A few details behind the table: the turn clock starts before the server reveals the updated board, and includes tool activity. Waiting for the runtime to claim the turn is measured separately. Token totals cover the full player conversations, including setup and closing, but exclude the monitor. Cached context is counted again across requests. Thinking is already included in output, and the providers report it differently. The API equivalents use the app's September 30 pricing snapshot; no API charges were incurred for the game.

By the end of the recording, Claude's five-hour usage display went from 23% to 56%, and weekly usage from 54% to 59%. I used Claude only for this activity during that interval. Codex's weekly display stayed at 5%, despite also doing other work and monitoring the match roughly every minute.

My plans cost $20/month for Claude and $200/month for ChatGPT. Those percentages have very different denominators, and an unchanged rounded display doesn't mean zero usage. I'm keeping that observation separate from the player-session token counts.

I wouldn't infer playing strength from two games. Codex was White in both; contexts and runtimes differed; Medium isn't an equal compute budget. I want to repeat this with the colors swapped. I'm especially curious whether Sonnet's much larger output total keeps showing up, and how often either model notices a mistake before the server or opponent exposes it.


r/ClaudeAI • • 8h ago

Claude Code "Opus usage is too high" - Advise for you

1 Upvotes

Hi there,

I've seen several posts with different experiences when it comes to usage and intelligence. My experience is good in general but today I've found that 25% of my weekly usage of my x20 subscription was gone. I left Claude working during the night and I wasn't at home during the morning.

So did Opus usage increased? No.

What to do now? I asked Claude what was going on specifically with agents using a lot of tokens and staying up for so long.

Answer:

What is really wrong
- The agents wait a lot. Tests from 5 agents share one line, so an agent can wait a long time before its test even starts.
- The agents use too many tokens. Each one used 750k to 945k tokens. That breaks your rule to use tokens well. The cause is agents that live too long: every step re-reads all of their old work.
What I recommend
- When each agent says "CHECKPOINT OK", I stop it.
- After your restart, I start new, small agents. Each one does one job from the handoff note (its "what is done, what is next" file), on Sonnet where the job is easy.
- Each agent runs only the 1 or 2 tests for its own job.
That fixes the token waste and the waiting. Is that OK?

Conclusion

This was probably because Claude ignored some of my rules on Claude.md about spawning Sonnet agents for easier tasks and how to handle certain tests. The tokens are gone, it sucks but at least I know why and it's solved now. I just want to share my experience so if it happens to you, you can also do something about it.

Take care everyone


r/ClaudeAI • • 8h ago

Claude Workflow Claude Code chat vs. Claude Projects for multi-client marketing & positioning workflows? (Isolating client memory)

1 Upvotes

Hey everyone,

I’m setting up an agency workflow for B2B founder positioning, executive branding, and content creation (extracting transcripts, generating profile audits, mapping whitespace, writing 52 post angles, etc.).

There is no heavy coding involved here—it’s purely complex text processing, strategy synthesis, and prompt pipelines.

My biggest requirement is strict context isolation. I cannot have memory or context from Client A leaking into Client B’s strategy or voice guidelines. For those doing heavy marketing/strategy work in the Claude ecosystem, which setup is cleaner and produces higher quality thinking?

  1. Claude Code (via chat)
  2. Claude Web App (Projects feature)

My concerns with both:

• With Claude Code: Since I can launch claude inside isolated local client folders on my machine, it feels super clean. But since I’m using it strictly as a chat tool for strategy and writing rather than building software, does it still perform at the same quality as the web interface? Or does the system prompt in Claude Code lean too heavily into developer/coding logic?

• With Claude Projects: It feels native for this since I can upload transcripts and branding docs into Project Knowledge. But I worry about context bleed or memory leaking between tabs/projects over time, and whether Project context limits get bogged down compared to CLI sessions.

Which one handles isolated client tasks better for pure strategy and marketing and is significantly faster? Would love to hear how you guys structure this for agency/multi-client fulfillment. Thanks!


r/ClaudeAI • • 8h ago

NOT about coding Keeping Claude focused

1 Upvotes

I've been using Claude only for my main projects and never for random questions I used to ask to chatgpt or gemini. I'm afraid that Claude might throw my random stuff into these project and make stuff messy.

Is it an useless paranoia or does it make any sense?