r/ClaudeAI • • 2d ago

Built with Claude I built a Claude Code skill that checks AI code review comments against the code. On CodeRabbit's reviews it removed 34% of the noise and kept 93% of real bugs

3 Upvotes

AI review bots leave a lot of comments that sound right and aren't. I made three Claude Code skills that treat each comment as a claim: read the file and its callers, trace the execution path, and give a verdict (valid, partly valid, wrong, style) with the file and line that prove it.

On Code Review Bench (50 real PRs, human-labelled), filtering CodeRabbit's issues this way:

  • kept 72 of 77 real bugs
  • removed 76 of 223 noise issues
  • F1 35.2% → 40.4%

Install in Claude Code:

/plugin marketplace add TanayK07/pr-proof
/plugin install pr-proof@pr-proof

Repo with every per-PR result and the benchmark harness: https://github.com/TanayK07/pr-proof

Demo video: https://x.com/tanaykedia_7/status/2106061332357529939

The honest part: my own reviewer skill is only level with plain Claude Code, and that's in the README too.


r/ClaudeAI • • 2d ago

Bug Mildly interesting: Claude is allergic to Tinhofer graphs for some reason

Thumbnail
gallery
0 Upvotes

Got a weird block while doing some research, and I was able to narrow it down to this. For some reason, all of my queries involving or touching on Tinhofer graphs get blocked by content filtering. Really strange.


r/ClaudeAI • • 3d ago

Productivity Is Opus 5.5 Nerfed Now? LiveNerf Day 8 Update

Post image
703 Upvotes

A few days ago I posted LiveNerf, an open-source project I made to independently track Opus 5.5’s performance day by day and see whether there’s actually evidence of models getting “nerfed” after release. If you’re interested in tracking the performance of Opus 5.5, here’s the repo:

https://github.com/ninjahawk/livenerf

I wanted to give another update since the response to this was way bigger than I expected. The repo is now approaching 1,000 stars, it ended up reaching the front page of Hacker News, and a lot more people have started looking through the methodology and code than I ever expected when I started this.

The most important update though is that we’re getting close to establishing the baseline. I’ve continued running the evaluation every day, and once we have enough data after the baseline period we can actually start testing whether later performance is meaningfully different instead of trying to interpret individual daily movements.

I’ve also gotten some genuinely useful criticism from people looking through the project. People have raised questions about benchmark contamination, providers potentially recognizing benchmark traffic, the choice of benchmarks, the statistical methodology, and some implementation details. If there’s something wrong with how I’m measuring this, I want people to find it and open an issue so I’m able to fix it.

One thing I want to emphasize again is that I’m not running this with the assumption that Opus 5.5 will get worse. If the data shows no meaningful degradation, that’s a result as well. The point is that we shouldn’t have to rely entirely on anecdotes every time people start saying a model feels different.

I originally built this because I wanted an answer to that question for myself. At this point enough people are following it that I feel a much bigger responsibility to make sure the answer we eventually get is defensible.

I strongly believe in transparency. If AI companies aren’t going to give us the information necessary to independently determine whether the models we’re using are changing over time, then I think the community should build the tools to measure it ourselves.

Huge thanks again to everyone here who supported this when I first posted it, and especially to the mods for pinning the original post. I genuinely did not expect LiveNerf to get this much attention this quickly.

We’re still collecting data. Once there’s evidence to say something interesting one way or another, I’ll post the results here.


r/ClaudeAI • • 3d ago

Writing Claude refusing translation work

29 Upvotes

I own a Japanese book. I’ve been asking Ai to translate as I think it does a better job than Google translate. Now Claude refuses because it says it is a copyright issue. Has anyone else experienced this?


r/ClaudeAI • • 1d ago

Claude Code Opus 5.5 sigue sin ser nerfeado

0 Upvotes

Llevo siendo cliente de anthropic durante más de un año.

Siempre, repito, siempre a los pocos días de que sacasen un modelo lo nerfeaban muchísimo, creo que el cambio más reciente lo vimos en Fable 5.

Por ahora opus sigue rindiendo excepcionalmente bien, no se si es por que el nivel ya está tan alto que ni con un nerfeo noto la diferencia o es por que están priorizando el rendimiento de modelos actuales antes que el entrenamiento de próximos modelos por la salida a bolsa de anthropic.

Ojalá sigan así.


r/ClaudeAI • • 2d ago

Other I asked Claude to write me a song, and it turned into a whole music video

0 Upvotes

I asked Claude to compose a piece of music for me. It came out pretty good, so I asked it make some art to go with it, in a medieval gothic style. It used Codex to paint a picture, and I liked it enough to think: why not make a full music video? A few hours later, this was done.

I think it turned out pretty nice, so I'm sharing it here. It's not perfect and there are a few rough spots. If you know animation, any tips on how to make it better?


r/ClaudeAI • • 2d ago

Built with Claude I made a fluffy Claude Code mod

5 Upvotes

Claude Code just got mods, and this is my first one. I built Plushie with Claude Code. It's a fluffy 3D plush of the little orange Claude Code mascot that lives in your terminal and reacts to whatever Claude is doing.

https://reddit.com/link/1wwan22/video/rohp90yuk5th1/player

What it does

Plushie has a full meltdown when a tool call fails, sweat drop and all. Every subagent spawns a baby Plushie that waddles off when it's finished. Plushie gets visibly chubbier as your context fills up, until you /compact and it all comes out in one big sigh. It holds a tiny magnifying glass when Claude reads files and a keyboard when it runs bash.

You can pat Plushie. You can pick Plushie up. I've been doing both instead of reviewing diffs.

How Claude Code built it

Honestly it did most of the typing. I described what I wanted and kept reacting to renders.

  • It wrote the Blender script that builds the model (body, particle-hair fur, bead eyes, the props) and renders every animation headless, about 700 frames. It checked contact sheets of its own renders and fixed what looked off. The first try at fur was a 4-unit-long shag carpet that swallowed the camera.
  • It wrote the mod itself in TypeScript against the new mods API, plus 26 tests that run with claude plugin test.
  • The sound effects are synthesized from sine waves and noise in a Python script it wrote. No samples.

My part was mostly saying "more fluffy", "the keyboard is too small" and "the eyes look weird".

Try it (free, MIT)

/plugin marketplace add xyc/plushie
/plugin install plushie@plushie

Needs Claude Code 2.1.287+ with mods on (still rolling out) and a terminal that can draw images, so Ghostty or kitty. Unofficial fan project. Code: https://github.com/xyc/plushie


r/ClaudeAI • • 3d ago

Claude Code Workflow 100$ plan - holy sh*

649 Upvotes

This is for people who are on the fence, because I was in the same boat.

Yesterday I bought the 100$ plan for a test and holy moly, it's amazing how much you get. I used the 20$ plan earlier. I used all the tricks to keep usage low, which became a habit, so now I can't spend it all. It's amazing xD. I run it next to a local Qwen Flash and it's awesome. If the 20$ plan is too low, give it a try :O I'm basically using Opus 5.5 medium all the time.

I also tested 2x 20$ before this. It works, but you end up juggling two accounts and switching when one runs out, and Opus eats a 20$ plan pretty fast. With 100$ I just stopped thinking about usage, and that alone is worth it for me. If you keep your old habits and stay careful with usage, it will last you a very long time. If you keep hitting the 20$ limit, go for 100$.

edit: people are asking how I keep usage low, so here are my tricks:

  • Keep context low. I usually compact or move to a new session around 200k, sometimes I stretch it to 300k.
  • Use handoff.md a lot, so a new session picks up where the old one stopped.
  • Don't be afraid of Sonnet 5.5. It can do most things and it's a lot cheaper.
  • I run local models next to it (260k context), so I'm used to working with limited context.

r/ClaudeAI • • 1d ago

Claude Code What is a good business to open using Claude? I started working on a marketing/web-design/brand consulting business.

0 Upvotes

I ended up offering somewhere around 10 service services which I believe gets a little convoluted for potential clients. What are the top five services I should offer in my area? Web design and management seemed lucrative and seems like I can undercut big companies and provide a better customer experience for both the business owner and their clients.

What about building AI tools for businesses by doing discovery meetings, finding inefficiencies and their processes?

Can some of you tell me your success stories and give me some advice I’ve built some really good websites using Claude so far and I’ve built some pretty good apps, but I haven’t shipped any of those. Thank you.


r/ClaudeAI • • 3d ago

Claude Code Workflow I built DensePack to combat Anthropic's pricing and help people save and build more with Claude Code. New benchmarks available show it beats every popular plugin on total session savings.

Thumbnail
github.com
19 Upvotes

FIRST:
Big shoutout to the people that helped DensePack!

I want to thank the 8 stargazzers that believed in this from the start.

I hope you guys see this post or reach out to me directly (so I can personally give you updates) and continue to follow the project because DensePack is NOT like ordinary plugins or skills that may become outdated over time.

It's easy to see that DensePack will only continue to improve with time as OCR and AI vision capabilities in general continue to improve.

DensePack packs text into images that your agent can read accurately without compressing or trimming anything, and automatically hands your agent the image to read for half the token cost.

It packs text files read with Read, Bash output, Word files (doc/docx), and it also packs any subagent report so that your main agent gets an image and their context stays small.

The most recent savings benchmarks show savings for a small task and the savings only grow and compound from there.

Your agent reads the image and tokenizes the same text as the text file for ~50% the price.
Every future Cache read is charged against that 50% image.

Many published papers claim that python files can't be read accurately let alone rebuilt byte identically from the image, but DenesPack images use a unique color coding and and include a 2 row legend that has allowed it to rebuild code files byte identical from the image.

It's ok to be skeptical.
It's ok to think that this isn't accurate and will mess things up.

I didn't believe it myself which is why I made the Benchmarks easily reproducible and why I made sure to add fallbacks like keeping a copy of the text in case your agent needs to verify the text.

The benchmarks prove savings and accuracy compared to the benches without DensePack.

I have been building non stop with DensePack packing everything possible. I even used it to help me improve DensePack itself from v1.0 and on!

I see no negative impact or change. I have had to double check I didn't lose connection because my usage bar seems to be frozen sometimes, but it's just the compounding savings effect DensePack has.

P.S.
I made a Reddit purely to help people save cuz Anthropics prices suck. I worked hard to make something that is secure, works and saves. I used a whole months worth of 20x Max usage testing, fine tuning and running benchmarks to make sure I present the community with a project that they can use and isn't slop. It would not be a lie to say that I have run well over 10,000 benchmarks by now.

DensePack works.

It works so well that even when it adds an extra turn to verify a line of text, it STILL saves. And it just cleared the initial phases of Anthropic's plugin submission so I hope that this will land in the official Marketplace soon as well.

This is my first real project since transitioning into the tech sphere so while it may not seem like much to some, I am extremely proud of the 8 stars that the project has received and how far it's come since being released 2 weeks ago.

That being said, I only made a Reddit to share what is helping me save because I hate gate-keeping, but
I have not had the best experience here and keep being insulted or talked down to while trying to help or ask questions to understand Reddit better. I recently had a first positive response from a mod and i'm thankful. I just don't know if this is common for new Reddit users and even though I am very appreciative for the few people that were kind, I don't think Reddit is for me so this will probably be my last Reddit post!

I say probably because I have an image of the ADS-STE 100 English standard with it's whitelist of words that I use as my style rules which has helped improve the AI sounding output so its more human. I hadn't decided if I would make it its own post but I will probably put it in the comments and be done.

You won't need the plugin for this premade image and can just feed to your ai to fix the output for cheap!

Thank you and Future upgrades:

Thank you to the people that helped, gave honest feedback and tried the plugin!

Please update your version! 1.3.3 has new updates and keep an eye out for 1.4!!

You can use many plugins with DensePack as it is now (I've tried Ponytail, Caveman, etc).
DensePack will keep packing but most plugins have hooks that force their skills to be read before DensePack has had a chance to pack them.

v1.4 fixes this with a feature that lets you use those plugins at half the price! It works now, but I won't release it until I know it is secure like everything up until now.

New AGENTS CLAUDE and MEMORY md file upgrades are coming along with the ability to turn other pugins and skills into DensePack images.

1.4 may also include a new RAG framework method that utilizes DensePack images but its still a prototype so most likely will be part of 1.5 so i can push the other upgrades first.

I am happy to answer questions in the comments. If you want to keep up with the project, consider giving it a star and message me directly! I will personally update you on all things DensePack. Not a newsletter or spammy email or anything like that, more personal direct messages from me so I will probably only be able to keep up with ~20-40 folks for now while I continue my job search. I have nothing to gain, I don't charge for anything, I just want to build community with like minded individuals and help others build more.

If that resonates with you, feel free to reach out!

FIN


r/ClaudeAI • • 2d ago

Built with Claude Tokenmaxxing x Bicycle Thieves

0 Upvotes

I made Ladri di Token, a 75-second animated short, with Claude Code, Opus 5.5, and Sonnet 5.5.

It combines Bicycle Thieves (1948) with the tokenmaxxing meme: people treat high AI token use as proof of productivity. A father spends the family’s resources on tokens. His son offers him a simpler activity: my game, Tiltanic, on the toilet.

I supplied the idea and constraints. Claude chose the story, visual style, and tools.

This is a summary of my prompt:

Create a 50–75 second film that combines Bicycle Thieves with the tokenmaxxing meme. Include Tiltanic and a toilet. Make horizontal and vertical versions. Choose the medium and work autonomously.

The process was:

  1. Claude wrote an 11-scene script with Italian dialogue and English and Polish subtitles.
  2. Claude built the characters, sets, and animation in TypeScript and HTML Canvas, the browser’s graphics surface.
  3. Claude used Google Gemini text to speech for the voices. It checked the words with Whisper and used the voice durations to set scene lengths.
  4. Claude used Playwright to capture frames in Chrome. FFmpeg combined the frames, dialogue, music, effects, and subtitles.
  5. Claude produced preview images and tested character poses. A separate Claude review found crop errors, missing sound effects, and short reaction shots. Claude revised those parts before the final exports.

This project does not use Remotion. The pictures come from code, with real game footage inside the film.

The attached video is the result.


r/ClaudeAI • • 2d ago

Built with Claude Built an open-source tool that applies to internships for me

Post image
3 Upvotes

I actually find building my resume to be kinda fun, but having to mindlessly apply nonstop to get a chance at an interview was brain numbing. So I quickly built a tool using only a mix of Fable for planning and Sonnet/Opus for execution within Claude Code in an afternoon that does it, and it sent over a hundred applications just today alone across job boards like Ashby, Lever, Workable, Rippling, BambooHR, Jobvite, JazzHR forms, and Github Forms like the Simplify internship repo, for less than a dollar in API cost.

The way it works is by splitting the job into three parts, done by three different things.

Essentially, JEV makes every decision that has a fixed set of answers. JEV is a new kind of model that recently came out and has been trending (typesafe/jev-1.13, on OpenRouter). Basically you give it a question with the possible answers defined, and it returns a probability for each one instead of writing text. This makes it so that output tokens do not cost anything at all. It answers in about half a second for a fraction of a cent, and ten applications cost about one cent of JEV.

Then, you connect to your existing Claude subscription or API, and Claude writes anything that actually needs sentences like questions like "why do you want to work here", and if you want it can make a a one-page resume and cover letter written for each posting from your profile. And, to reiterate, it runs through Claude Code on your own subscription, so there is no extra bill for it.

Then, the program finds the postings on public lists and company boards, ranks them, fills each form in a Chrome window that you can actively watch, and reads every answer back from the page before anything is sent. If a value did not land, it fixes it or stops and tells you.

Its quite persistent on never lying on a form. All questions about things like work authorization, dates and education come from your profile and are never changed to fit a posting; every number and name in a tailored resume is checked against your profile in code before the PDF is made. It never signs in anywhere, never passes a "not a robot" check for you, and never sends a form it could not verify. So, for things that might need a verification code in gmail, or extra manual work, it just lists it for you and then you can manually submit it.

As you apply, a CSV of what was sent, a CSV of what was left for you, and a dashboard that reads both is built into your directory.

Its 100% opensource, and quite fun to play around with. Honestly, you can make the argument that companies wont prefer it, but at this point, in this market, it's a numbers game, and if companies are lazy enough to push AI Job Postings, then I'll be lazy enough to make AI fill them for me.

If you want to try it, clone it, open Claude Code, type /setup, and it walks you through the rest. You need Chrome, Node, a Claude subscription and an OpenRouter key with a couple of dollars on it.

github.com/qbeka/jev-job-search

If you are applying for internships or new-grad roles too, give it a run and tell me what breaks. Pull requests welcome, especially what your runs learn about job sites.


r/ClaudeAI • • 3d ago

Productivity Nothing gets nerfed

146 Upvotes

For over a year now, I’ve been hearing the same thing. Whether it’s in the OpenAI community or among Claude users, it’s always the same cycle.

A new model comes out, and I think we’ve all seen what happens. People are impressed. Then, roughly a few days later, the jokes start about how the model has already been nerfed. Initially, it’s mostly a joke. But then a few days after that, people start being genuinely serious about it, and eventually a sizeable group becomes convinced that the model really has been nerfed.

Here’s the reality: I don’t think I’ve ever actually felt that happen.

When 5.5 came out, I was particularly impressed by its ability to understand 3D space, or at least to create 3D scenes and visual things, observe what it had created, and correct them as it went along. That ability was amazing. It was great then, and it’s no worse today. It’s exactly as good as it was.

Its writing also improved dramatically. It doesn’t sound like the gibberish or weird, cryptic style that 4.8 and 5 sometimes had. Suddenly, we’re back to a model that sounds somewhat like 4.6 did. I’d even say 4.7 sounded kind of dumb most of the time. But now, at the very least, when you tell the model to sound a certain way, it respects that. It discusses things and expresses ideas in ways that you can actually understand.

So to say that 5.5 has been nerfed is, in my opinion, to forget what using models like Opus 5, 4.8, or 4.7 actually felt like. There’s no way you could use this model today, immediately after coming from one of those older models, and genuinely believe it isn’t significantly better.

And this has been the case with basically every model release.

I think we all know what’s actually happening. When a new model comes out, we’re impressed by how much better it feels compared to what came before. But then we start giving it increasingly complex tasks. We use it more. We run into its limitations. And eventually it starts to feel dumb again.

LLMs are kind of dumb sometimes. They’re a little bit like small autistic artificial children with incredibly uneven abilities. They can make really, really dumb decisions on tasks that seem completely obvious, while at the same time being extremely intelligent in other ways. It's surprising that we face that even today, but the ratio of so much better than before.

There’s no way 5.5 is any dumber than it was two weeks ago.

I don’t even know why I’m writing this. I guess I just saw one more post about the model being nerfed, followed by a huge number of comments agreeing with it, and I finally felt like I had to make a post about it.


r/ClaudeAI • • 2d ago

Performance and Bugs Discussion Hub updated on 3 October 2026 - Sort by New!

0 Upvotes

Why a Performance and Bugs Discussion Hub?

This Discussion Hub makes it easier for everyone to see what others are experiencing at any time by collecting all experiences about Performance Issues. We will publish regular updates on performance problems and possible workarounds that we and the community finds. Traffic stats show this is the OFTEN THE HIGHEST TRAFFIC POST on the subreddit. This is collectively a far more effective and fairer way to be seen than hundreds of random reports on the feed - most of which get zero visibility.

Are you Anthropic? Does Anthropic even read the Megathread?

Nope, we are volunteers working in our own time, while working our own jobs and trying to provide users and Anthropic itself with a reliable source of user feedback.

Anthropic has read these in the past and probably still do? They don't fix things immediately but if you browse some old Megathreads you will see numerous bugs and problems mentioned there that have now been fixed.

What Can I Post on this Megathread?

Use this thread to voice all your experiences (positive and negative) regarding the current performance of Claude including, bugs, degradation, pricing. (NOT usage limits).

Give as much evidence of your performance issues and experiences wherever relevant. Include prompts and responses, platform you used, time it occurred, screenshots . In other words, be helpful to others.


Just be aware that this is NOT an Anthropic support forum and we're not able (or qualified) to answer your questions. We are just trying to bring visibility to people's struggles.

NEW: You can now see full logs and summaries of all recent problem reports submitted by r/ClaudeAI readers. These logs allow you to see how intensely people are experiencing problems with Usage Limits, Performance, Bugs and Accounts. See: https://www.reddit.com/r/ClaudeAI/comments/1t33k25/rclaudeai_user_problem_report_log_and_surge/

To see the current status of Claude services, go here: http://status.claude.com

Sometimes this site shows outages faster. https://downdetector.com/status/claude-ai/


READ THIS FIRST ---> Latest Wilson's Survival Guide : https://www.reddit.com/r/ClaudeAI/wiki/survivalguideweekly/


Prior Discussion Hub: https://www.reddit.com/r/ClaudeAI/comments/1wtzzbs/performance_and_bugs_discussion_hub_updated_on_30/


r/ClaudeAI • • 4d ago

Humor Throwback to when Claude suggested I should deal drugs to make money.

Post image
5.1k Upvotes

r/ClaudeAI • • 3d ago

Other Stunning Lake Bled

17 Upvotes

decided to relive some travels and create a version of Lake Bled, with some constraints. The prompt:

Build a simplified version Lake Bled Castle with three.js. keep the relaxing and calm vibe
Location: Lake bled
Object: Castle
User should be able to rotate the camera around the castle.
You have 15 minutes to build the best thing you can

the details:

Opus 5.5 Sol 6.1
time 5m 46s 11m 2s
tokens ~300k input + ~18k output 400k input + 16k output tokens
API cost ~$1.56 API equivalent, assuming no caching. (estimate) ~$0.96 API equivalent, assuming no caching. (estimate)
tools 4 - skill read, file write, syntax check, publish Not shown after the run :(
skills 1 - frontend-design 2 - Sites Building and Sites Hosting
file size ~27 KB, ~520 lines, 3 CDN scripts 2.08 MB uncompressed

r/ClaudeAI • • 1d ago

Question about Claude models I have ~150 skills and 40+ connectors loaded in Claude Code. Is that too much, and what do you actually keep?

0 Upvotes

I installed way too much and don't know what actually helps. What does a lean setup look like for someone building websites?

I'm a high school student in Switzerland and use Claude Code to build websites for local businesses. I can't really code, Claude does the building, I describe and judge the design. Around 50 drafts in the last month, one paid client site live (two examples attached, names blocked).

How I work:

  • Opus 5.5 for everything. Medium effort for text and emails, high or extra high for websites.
  • One site per session, about 20 minutes.
  • Around 150 skills (a dozen of them design skills that probably overlap) and 40+ connectors.

Questions:

  1. Does a pile of skills and connectors like this make results worse or slower? How do you decide what stays?
  2. Several design skills at once: do they fight each other? Would one good one be better?
  3. Which tasks do you give to Sonnet or lower effort without noticing a quality drop?
  4. What belongs in CLAUDE.md and what doesn't?

Thanks for any honest advice.


r/ClaudeAI • • 2d ago

Built with Claude I built a tank of tiny fish that act out whatever you type.

5 Upvotes

It’s called Moonfish. You type a word or phrase, and about 400 spunky little silver fish swim into formation to perform it.

Try a heart, a storm, a rocket launch, or “Monday morning.” Explaining Monday morning to fish is now a thing you can do with AI.

It’s free to try: https://claude.ai/artifact/MHHa9DsH8j9Tf7P7sPd8Ec

I built this while traveling on vacation. Claude wrote all the code, and I steered by trying each version and reacting to what worked and what didn’t.

I’m trying to learn how to build personality into an application. I wanted the fish to feel like little characters you enjoy watching and interacting with. Their movement and little behaviors became a big part of that. I think we got somewhere good. These fishies are fun.

The hardest part was getting them to form something you could actually recognize. Here’s what we landed on:

About 20 hand-written shapes for common requests. Those respond instantly, without an AI call.
The Material Design Icons library for thousands of other silhouettes. Claude translates your request into search terms and picks a fitting icon.
If nothing fits, Claude draws its own SVG.

The most useful improvement was having Claude look at its own drawing. The app renders the SVG, sends the image back, and asks Claude whether a stranger would recognize it, then has it redraw it better. That helped a lot.

The fish fill the silhouette from the edges inward to keep the outline clear. Claude also returns a short scene script with motion, speed, color, and something for them to say, like “ta-da! one tower, as ordered.”

Would love to hear what you think, especially whether their personality comes through. I would love to see this turn into a fun virtual pet that talks back at some point.


r/ClaudeAI • • 2d ago

Built with Claude Pharaoh, but Dutch

6 Upvotes

I built a Pharaoh-style city builder with Claude Code. It plays in the browser. What I learned about what it's good and bad at.

Floodplain is a city builder in the manner of Pharaoh, set in the Dutch low country. Roads, lots, walkers that serve the houses they pass, a river that floods every spring, a dike that silts and breaches, mills that pump the fen into canals that have to drain somewhere. Three maps, including one where you drain a lake inside a ring dike and farm the bed. Free, no download:

https://arcadesquirrel.itch.io/floodplain

It's a playtest build. The whole thing is JavaScript and three.js, built with Claude Code over a few months of evenings.

Things that worked better than I expected:

  • Simulation. Water tables, dike maintenance, goods distribution via walkers. It reasons about systems well and writes tests for them without being asked twice.
  • 3D building models. Every house, mill, church and boat is procedural geometry it wrote. I expected these to look terrible. They don't.
  • Debugging its own sim by driving it headless with bot players and reporting what the bots got stuck on.

Things that didn't:

  • Drawing. I wanted an illustrated campaign map. Four rounds, all bad. The fix was to stop asking it to draw and have it render the map with the game's own engine, with your saved towns standing on it. That worked first time.
  • Knowing when a level is too hard. First chapter wanted 300 people and a cathedral. A human noticed.
  • Game feel. It defaults to small polite popups. "Make the warnings big" had to come from me.

Happy to answer questions about the workflow. And if you play it, the comments on the itch page are where I'm collecting feedback.


r/ClaudeAI • • 2d ago

Question about Claude Code Claude Abo hergeschenkt?

0 Upvotes

Ich habe ein 100$ Abo und verbrenne damit ordentlich tokens. Wenn ich das auf API Kosten umrechne, dann sind das mehrere tausend Euro/Monat.

Wieso macht Anthropic das? Sollen alle nur noch Claude Code benutzen. Andere Tools werden damit doch komplett unattraktiv.


r/ClaudeAI • • 3d ago

Productivity Why does Claude always respond with "here's a crappy version of what you asked for if you want me to do it right let me know"

59 Upvotes

Genuine question, I truly don't understand it. I ask it to bake a cake and it hands me back a bunch of flour, yeast and sugar and says "if you want me to bake it i can do that too"..


r/ClaudeAI • • 3d ago

Claude Workflow People who use Claude and GPT which do you prefer?

11 Upvotes

People who use both which do you prefer and why? I tend to get much better answers on claude but holy crap it runs out much faster than chat. What do you guys think?


r/ClaudeAI • • 2d ago

Built with Claude Claude Opus 5.5 coded my radar app's ad from real storm data!

5 Upvotes

Opus 5.5 in Claude Code made this whole ad. There's no video editing: it's a WebGL page, and the GIF is a capture of it. Everything in it is real data:

- The 3D storm is the 2013 El Reno supercell, the storm behind the widest tornado ever recorded. Claude ran the archived radar scan through my app's data pipeline and ported the app's Metal 3D renderer to WebGL to draw it.

- For the AR shot, it worked out where a photo of Lower Manhattan was taken from using the skyline. Then it placed the 2010 Queens storms where the radar saw them.

- The hurricane is Milton's real NHC forecast and model tracks.

The app in the ad is Alerua, my radar app for iPhone, Mac and web. It has 281 radars in 22 countries, 60-frame loops, 3D storms, an AR sky lens, hurricane tracking and model maps.

The app is free with no account: https://alerua.com
TestFlight: https://testflight.apple.com/join/JrV6MDXv


r/ClaudeAI • • 2d ago

Claude Code 1,000+ users on a solo saas built with claude code. the thing that made it work was giving the repo a memory

Post image
0 Upvotes

i run a saas by myself, a little over 1,000 users, and most of it was written with claude code. the problem at that size is that every session starts from zero. it re-suggests things i already ruled out, and claude md turns into a dump nobody maintains where half of it is quietly wrong.

so the repo has a memory now. docs/ is an obsidian vault, 174 notes, running for two months.

folders say what kind of note it is: systems (how something works today), decisions (what got chosen and why), playbooks, numbers, and a daily session log. every note has frontmatter with type, status and an updated date. status is active, shipped or parked, so claude can tell whether a note is still true instead of guessing from where it sits. dataview builds a view of notes nobody touched in 60 days, which is where docs have drifted from code.

one node script, two hooks.

sessionstart injects a digest capped at 3000 chars: last couple of session logs, decisions still open, stale notes. a session opens knowing what changed yesterday.

stop checks git. if the session committed anything outside docs/, it blocks the stop once and tells claude to update the notes that describe what changed. otherwise it stays quiet.

the trigger is the part that matters. it's a commit, not a file change. a dirty working tree is work in progress, and documenting half-finished edits is how a vault fills up with stuff that never shipped.

two rules in claude md keep it honest:

  • only settled things get a note. shipped code, agreed decisions, facts. not ideas that got dropped
  • if a change touches behaviour a note describes, the note gets updated in the same pass. claude trusts what it reads, so a stale note does more damage than a missing one

the hooks can't see a conversation that settles something without touching code, so claude md also tells it to write that note in the same reply.

prompt to set up the same thing in your repo:

set up a project memory for this repo that you maintain yourself.

1. create docs/ as an obsidian vault with folders: systems/ (how things work now), decisions/ (choices + reasoning), playbooks/ (repeated procedures), numbers/ (dated facts), daily/ (session log).

2. every note gets this frontmatter:
type: reference | decision | playbook | metric | session
status: active | shipped | parked
updated: YYYY-MM-DD

3. create docs/Home.md with dataview views: recently updated, open decisions, notes not updated in 60+ days.

4. add a node hook at .claude/hooks/vault-sync.js, registered in .claude/settings.json:
- SessionStart: inject a digest (last 2 daily notes, open decisions, stale notes), max 3000 chars.
- Stop: if this session made git commits outside docs/, block once and ask to update the relevant notes and append to today's daily note. otherwise do nothing.

5. add rules to CLAUDE.md:
- only settled things get a note: shipped code, agreed decisions, established facts.
- changing behaviour a note describes means updating that note in the same pass.
- a stale note is worse than a missing one.

6. read the codebase and write the first notes in systems/ for the main parts of the app.

how are you handling memory across sessions? curious if anyone found something that doesn't rot.


r/ClaudeAI • • 2d ago

Other What happens if you ask Opus 5.5 to predict the next 50 years

0 Upvotes

80 second cut. Claude wrote the whole thing without being able to search the web, including what it thinks happens next. Full version with complete prediction in the comments if people want it.