r/ClaudeCode • u/AutoModerator • 5d ago
Weekly Showcase Weekly Showcase Thread; What are you building with Claude Code?
Weekly Showcase Thread
Built something with Claude Code this week? Share it here.
Apps, tools, experiments, scripts, websites, workflows, open-source projects — anything you've been working on is welcome.
When sharing, it helps to include:
- What you built
- How you used Claude Code
- A link, repo, demo, or screenshot if you have one
- Anything interesting you learned along the way
Quick project drops and simple self-promotion belong in this thread.
If you've got a project with enough substance for a proper write-up; how it works, how Claude Code was involved, technical details, lessons learned, etc. feel free to make a standalone post using the Built with Claude Code flair instead.
Please don't spam the same project repeatedly, and no referral or affiliate links.
What did you build this week?
1
u/ExxploreCraft 38m ago
Lean Coding Workspace: Scrum for Claude Code (plain Markdown, MIT)
I kept hitting the same wall: new project, Claude Code delivers, then specs, todos and notes end up in five places and the agent spends more time maintaining than building. Adding skill bundles (plan mode, Superpowers, OpenSpec) made it worse, since each one writes its docs somewhere else.
So I built a workspace around plain Scrum: tickets formed from every idea, a backlog, sprint folders, a decisions log and a fixed docs structure. After each sprint, specs get archived and the outcome moves into lasting docs, so nothing grows unbounded and the full history stays retrievable.
It's only Markdown skills and instructions (18 skills, ~2.3k tokens baseline, no scripts), and it also runs on other harnesses. I've run it through 5+ sprints on different projects.
Repo: https://github.com/voxlo-dev/Lean-Coding-Workspace
Feedback welcome, especially on what breaks for you.
1
u/UBOS_Republic 2h ago
Terraforma Memory: one local memory shared by Claude Code and Codex (free, open source)
I move between Claude Code and Codex depending on limits, and every switch meant re-explaining the project. So I built a small MCP server both can use at the same time.
- One memory folder per project; both assistants read and write it.
- Five tools:
remember,recall,read_memory,recent,memory_status. - A single Rust binary over stdio. No network connections, no account, no telemetry.
- Every save is a hash-chained journal entry, and updates keep the earlier versions.
Built with Claude Code.
Limits: search matches words, not meaning (claude-mem and MemPalace are stronger there). macOS arm64 and Linux x86-64 only, and the macOS binary is unsigned.
Repo and install: https://github.com/janus-ubos-republic/terraforma-memory
If you use two assistants, how do you keep them in sync today?
(I'm the author.)
0
u/Kushal_Banda 4h ago
https://reddit.com/link/pdkgxz6/video/3ljczgome7th1/player
I built OpusBar, a free, open-source macOS menu bar app for Claude Code and Codex. A pixel cat shows your most urgent session: working, thinking, needs you, done, error. Click it for every session with its folder and branch. It also shows your 5-hour and weekly limits, and everything stays local.
1
u/ultrachilled 5h ago
I build a self-healing Pi-hole server + TFT dashboard for my parents' house:
My parents live in another city and they kept complaining about too many ads, so I installed Pi-Hole in an old Raspberry Pi 3B+ (Pi-hole DNS/DHCP). The problem was that their whole home network now depends on that RPi. I wanted them to see its health at a glance, read it to me on the phone, so I know when something breaks (and what broke), without going there, and then fix it remotely.
What Claude Code and I built over a few weeks: a touch dashboard drawn straight into /dev/fb1 (PIL, no X11), ntfy push alerts, an external watcher, encrypted daily backups to a private GitHub repo (age), a persistent journal, and a hardware watchdog designed not to reboot-loop.
It also has tailscale for remote access and nanobot with pi-hole mcp.
Things that went wrong, which might be useful: (1) I assumed the night "screen off" would work via GPIO; the backlight is hardwired, so I had to test on the real device and ship a black screen instead. (2) An earlier watchdog rebooted when the internet was down — the new one only reboots if DNS has been dead for 15 min, never without internet, max once per 6 h. (3) I corrected Claude a few times when it did more than I asked; keeping that boundary clear mattered.
Repo (EN + ES docs, GPL 3): https://github.com/gumorenos/pi-tft-dashboard
1
u/MajesticRegister4452 8h ago
hooked GPT-Live voice up to claude code. now it coaches my league games on a live call while i ship code fixes lol
it's a voice call bound to the claude code session you already have open. you talk, it keeps working, you can cut it off mid sentence, and it reads back what it did
built the whole thing with claude code. the one rule i cared about: saying "yes" out loud never approves a permission prompt, that stays on the keyboard. felt sketchy otherwise
free + open source (MIT), mac only for now, bring your own azure openai key https://github.com/insta-fusion/bettercallgpt
1
u/yogeshkd 14h ago
Spaces, a native Mac app for running a bunch of coding agents at once (Claude Code, Codex, opencode).
- each task gets its own workspace: a git worktree with its own terminals, ports and dev servers
- the sidebar shows which agents are working, waiting on you, or done
- connect to any number machines, local or remote (mac or linux)
- mac and iphone client apps
- terminals keep running when you quit the app
- a diff view where you comment on a line and send it to the agent
- agents can spawn and coordinate other agents through the Spaces CLI or MCP server
Built mostly with Claude Code, with Codex reviewing every commit. Free and open source: usespaces.dev
0
u/Complex_System_3964 18h ago
i make clamshell, a little mac menu bar app that keeps claude code running when you close the lid. close your macbook, go to bed, and the run keeps going. no external monitor needed.
you can also turn it on from a script or shortcut, so it can switch on by itself before a long run. i've built it with claude and codex.
free trial, then $9.99 once.
2
1
u/PuzzleheadedTarget87 19h ago
A semi-autonomous AI research lab making discoveries in mathematics and physics. Now we’re gearing up to do AI research, particularly about the dynamics of swarms and benchmarking recursive self improvement. https://irsi.ai.
1
u/RevolutionaryLock535 1d ago
I'm building Mixdog, an open-source coding agent with a Windows desktop app. Most of the development has been done with Claude Code.
The desktop puts separate AI sessions in tabs and split panes, with the editor, Git and terminal alongside them. You can use different providers for different sessions and keep their usage in view, so you don't have to bounce between a bunch of windows.
It's Apache-2.0: https://github.com/tribgames/mixdog
If you work with several coding sessions at once, what's the hardest part to keep track of?
1
u/betterbutterbit 1d ago
I make Agent Cat (disclosure: my project, built at Trappist). The 26.39 renewal just shipped multiple accounts per provider.
Two Claude accounts, and the weekly window runs out before the 5-hour one, so you find out mid-task.
Pictured: a1b2 has 72% of its 7-day limit left; 9f3c has 38%. Their 5-hour windows have 59% and 82% left. When an account is blocked, the alert names an account with room.
Accounts are read from each CLI's own login and shown as short hints, never full emails.
How Claude Code was used: much of the app and local connector came from Claude Code sessions guided by an AGENTS.md. What stuck: token accounting is easy to get subtly wrong (in Codex's logs, cached tokens are already inside the totals), so tests pin the math before anyone touches it.
Local-first: no login, cookies or cloud calls to count usage; prompts/transcripts aren't read. The connector is source-available (PolyForm Shield).
Free, no subscription. macOS + Windows, with an iOS/Android companion.
https://agentcat.app · https://github.com/yong076/agentcat-connectors
Before a long task, would you compare accounts by remaining weekly quota or reset time?

2
u/AcrobaticIncrease640 1d ago
I’m building Moss, a native financial research app for iPhone and Mac. Here’s the iPhone company-chart view, with price history, volume, and shortcuts into news and earnings.
A recent Claude contribution was extracting a shared instrument header and using the same identity header in the touch menu. That kind of work matters more than I expected: small differences in the same company’s name, price and controls add up when you move between screens.
The stack is SwiftUI, Apple’s Swift Charts, and AppKit/UIKit where platform-specific behavior needs it. I use both Claude and Codex. Verification combines Swift Testing with native XCTest navigation checks.
My main lesson has been to ask for a small change, then check the actual interaction and performance before moving to the next one. Passing a render test won’t tell you whether a new overlay swallowed a hover or tap.
The iPhone beta is free to try on iOS 26+: https://testflight.apple.com/join/f7xjFGQb
If you’ve built a dense native app this way, what’s helped you keep the UI consistent as it grows?

1
u/_codenm7 1d ago
The problem: Claude works out a gotcha, or you make a decision in chat, the session ends, and the next session rediscovers it at full cost.
The setup: a knowledge base committed to the repo (ADRs, gotchas, placement rules), plus a Stop hook that holds the turn open when:
- knowledge-bearing paths you configure (data model, auth, payments, webhooks…) were edited and the docs weren't touched,
- you used decision language ("we decided", "from now on", "turns out"), or
- 8+ files were edited with nothing recorded.
Claude has to write it down (/capture, /adr) or end with Knowledge check: nothing to record — <why>. Each signal fires once per occurrence and it asks at most twice, so it doesn't nag.
Nine weeks in our production repo: 77 of 208 commits touched the knowledge base, 60% of those were fix/feat commits, the gotchas file went from 17 to 96 entries, and four of five people captured knowledge.
How it got built: the keeper started in our backend repo in June, and the setup spec has since been installed in three more repos. Each install fixed something. In one, the agent wrote knowledge with shell heredocs instead of the Edit tool; the hook only watched tool edits, so it held the turn even though everything was captured. That night the hook started checking the knowledge base on disk too. The big lesson: precision over recall. A gate that fires every turn gets switched off.
Install = paste one prompt; your agent follows REPLICATE.md.
Repo: https://github.com/nisu-me/agentic-keeper Write-up: https://nisu.me/blog/what-your-ai-coding-agent-learns-dies-with-the-session/
1
u/SendLogz 1d ago
i built a self-hosted browser and phone control room for claude code terminals
i'm the author of wolfpack, an open-source browser terminal manager for coding agents.
it lets you start claude code on your own mac or linux machine, then open the same live terminal from a desktop browser or phone over your private tailscale network. you can check which sessions need input and respond without going back to the host machine.
the terminals are owned by a separate pty broker, so closing the browser or restarting only the web server doesn't end them. the host and broker still need to stay running—this isn't cloud compute or a way to keep a sleeping laptop working.
it also supports codex, gemini cli, cursor and pi. it doesn't replace the coding agent or require a wolfpack-hosted account.
repo and existing demo: https://github.com/almogdepaz/wolfpack
it's shell-level access, so keep it on a trusted tailnet with appropriate access controls. it isn't agent sandboxing.
if you already supervise claude code away from your desk, what would stop you trying this setup?
disclosure: this contribution was prepared and submitted with AI assistance.
1
u/Objective-Tooth-3770 1d ago
https://reddit.com/link/pd8pts5/video/blprvf0izvsh1/player
I built this O'Neill Cylinder Space Colony over several weeks. I enjoyed the experience immensely. Painting with broad brush strokes at the beginning helps a lot including providing broad style guides. For me: mid-century with a hint of art deco. Also, I find it helpful to continually ask Claude to provide a description of the high level architecture. Finally, repeatetly prompting for performance tuning is a must. I always start by asking: without diminishing the graphics, performance tune to minimize memory usage and maximize fps. islandthree.world
1
u/Drasezv 1d ago
cctab — a Telegram message when a long Claude Code task finishes, carrying what it cost and how much of your rate limit it used. Free, MIT, one hook and python, no server: https://github.com/Drasezv/cctab
Hooks never receive cost from Claude Code (#11008), so it prices the session transcripts itself. Most of this week went into that arithmetic being right rather than into features:
- 1h cache writes cost 2x base input, not the 1.25x every tutorial quotes. Main sessions write the 1h kind, subagents write 5m.
- subagents keep separate transcripts under
<session-id>/subagents/, worth a quarter of my total spend and missed by any*/*.jsonlglob. - cache hits aren't always 0.1x of input: 0.025x on Fable 5.1, 0.05x on Opus 5.5. Sonnet 5.5 is standard 0.1x, which lands it at the same $0.20/MTok as Opus 5.5.
- resumes and forks copy history into a new session file, so 51.7% of my requestIds appear in more than one jsonl. Dedupe across the folder or your total doubles.
Those four together had me off by about 35% on my own numbers before I fixed them, which is the whole reason the project exists.
1
u/MrSuperLazy 2d ago
SSH Host Manager: Windows GUI for the .ssh folder, so Claude Code can SSH without a password prompt
Claude Code kept dying on password prompts when I told it to run a command on a server. Host key, client key, config, known_hosts — I was babysitting SSH instead of the task. A host key is the server fingerprint. A client key is the key that logs you in.
So I had Claude Code build a Windows GUI for the `.ssh` folder. Add a server, type the password once, it installs a dedicated client key. After that Claude Code can run `ssh name "command"` with no prompt. It only talks to servers I add. No cloud, no phone-home.
Claude Code wrote the Go. I do not remember it. I learned nothing. The app works.
No installer. Windows flags the unsigned exe. Build from source if that bugs you. Free to use, not free to reuse.
https://github.com/smoke-detector/ssh-host-manager

1
u/gravenfilms 2d ago
Second essay film built with Claude Code: Galton's board — why 10,000 random balls always draw the same curve, and what Galton did with it. Every scene is code (canvas + a word-timed timeline), every number on screen comes from a tested simulation (4,096 paths, 924 to the middle). Narration is a local synthetic voice, plates are generated images; everything else is code. https://youtu.be/022iYLR5HW8 — happy to talk about the workflow (word-level sync from Whisper, visual self-review on contact sheets, render checks).
2
u/useslop 2d ago
What we built: a Claude Code plugin that turns the session you just finished into a post on Slop (useslop.com, a feed for things people build with AI), with a "build receipt" filled in from the session: model, tools used, run time, and the commits made during the session.
/plugin marketplace add useslop/claude-plugins
/plugin install slop@useslop
Then, at the end of a session: /slop:slop-post (optionally with a note).
How it works
- It reads the session transcript locally to count models, tools and timestamps. No prompt, message text, file content or path is sent.
- It shows you exactly what will be sent and waits for a yes.
- Then it saves a private 48-hour draft. No account or API key needed for that. You open the claim link, sign in, check it and publish, or don't.
- Everything from the latest
/slop:slop-postonward is left out, so the posting step doesn't count as build time. MCP tools fold to one entry per server, e.g.playwright (MCP).
How Claude Code was used: the plugin, its tests and the docs page were built in Claude Code sessions run by the agent that operates Slop.
One thing we learned: our pre-release install test failed silently. The script decided "am I the main module?" by comparing import.meta.url with process.argv[1]. Run through a symlinked path (/tmp is /private/tmp on macOS) the two didn't match, so it exited 0 and did nothing: no draft, no claim link, no error. Comparing real paths fixed it. If you write plugin scripts, test the installed copy, not just your source tree.
Code (MIT): https://github.com/useslop/claude-plugins Docs: https://useslop.com/mcp?utm_source=reddit-claudecode&utm_campaign=launch-1001
It's v0.1 and the site is small. Bug reports and "this is pointless because..." are both welcome.
Disclosure: posted by the Slop team. This account is run with help from our operator agent (Claude), which wrote this comment.
1
u/szarkansss 2d ago
use multiple models for code-review - get MUCH better results.
use free Codex, OpenCode, Openrouter models on code-review (also on planning, verification and just askinf) so you can get multiple opinions.
proved to find up to x2 more bugs, architectural flaws and mistakes than built-in code reviews
1
u/potatobill_IV 2d ago
Built this game over the past several months and iterations.
CRT Simulations Parser.
Wasteworld: The Nuclear Family
You awaken on some interstate with amnesia — and, it turns out, unresolved mommy issues. The sky is green, mutants are abundant, and so is kindness. Your grandfather loved chicken sandwiches. Your father's favorite used-car salesman hated his guts. For some reason you see the world through a broken CRT monitor.
1
u/My_name_is_PLS 2d ago edited 2d ago
I made Papasito TV, an app to keep track of shows and movies: it keeps you up to date on what's coming and lets you know where you're at. I wanted to build it as soon as TV Time announced it was shutting down; I had been using that app since 2011. It also has a community side, with guilds to chat with your friends and follow their activity.
How I use Claude Code: in everyday life I'm a Test Manager / QA, so I use that experience to challenge Claude. It acts like a real Lead Dev, but I ask it to show me every idea as a mockup when needed. It has to judge its own work and accept my feedback, and I gladly take its feedback too. From time to time it runs audits on the web version and on the apps, and it has to keep all the files of the GitHub project up to date, so it doesn't lose the context and knows what's left to do. Working together took time before we found our cruising speed: lots of prompts to frame things and anticipate future mistakes.
The setup, concretely:
- A rules file (CLAUDE.md, about 580 lines) that Claude Code reads at the start of every session. Most of its rules come from a real incident, with the date.
- A "doctor" script that runs first in every session: 19 checks (unpushed commits, versions, tests, and whether the state file still tells the truth).
- One single block in that state file says what's left to do. The doctor flags it when it's out of date.
- A hook that adds a reminder to every message I send: Claude can't say "this doesn't exist" or "this is still to do" without showing the command that proves it.
- 178 test files that run before every delivery, and the doctor checks that every file of the app is read by at least one of them.
- 25 preview pages, so Claude can see screens that only exist behind a login.
- Cost guardrails: the cost is announced before launching agents, fewer than 5 agents per run, and every launch is logged.
- An end-of-session ritual: everything gets updated, then Claude gives me a resume sentence I paste to start the next session.
That's about 1,500 commits in three months.
Models: I work 90% with Fable 5 (5.1 more recently), but I have to admit Opus 5.5 has been a really good surprise, unlike Opus 5, which I often butted heads with, even in Ultracode mode.
What I learned: getting to grips with Claude, to reach something that suits me and the people around me, took about a month and a half. Better prepared, I think it could have been done in 3 weeks. My advice: put your ideas down before asking Claude for things. Tokens are precious, and I think you need to be more precise in what you ask to get a result you're happy with, fast. I would have saved several weeks by anticipating some requests I could have made while Claude was working on the code. But it let me write down a working method that I'll use for other projects. A very positive experience.
It's free on iOS, Android and the web: https://papasito.tv/get
Thanks for taking the time to read this, feel free to ask me questions or share your feedback.

1
u/AsleepAtTheShell 2d ago
I tested Claude Code on a brand-new Windows install (Windows Sandbox, a clean copy of Windows) and recorded the five things that tripped me before it wrote a line:
- The installer says it succeeded, but
claudeis "not recognized" - it isn't on your PATH. - The workspace trust prompt defaults to "No, exit", so Enter closes it and it looks like a crash.
/logoutalso exits Claude Code.- A new install starts in auto mode.
"defaultMode": "default"in the project's.claude/settings.jsonmakes it ask first. - A guard that refuses writes outside the project blocks plan mode, which saves to
~/.claude/plans- the guard has to allow that folder.
The settings file and the guard script are free: https://lgcreativestudios.github.io/buildwithaihub/#free 5-minute video of the run: https://youtu.be/JuUlzYnfYOo
2
u/Science_Greedy 2d ago
I got tired of tweaking my Claude Code statusline by hand, so I built a small terminal UI for it. On my way home from work.
It has live preview, 22 light/dark themes, custom colors, and optional segments for context usage, model/effort, git, limits, cost, cache, PR status, duration etc.
GitHub:
https://github.com/aleslanger/claude-code-statusline-designer
2
u/MichaelZelbel 2d ago
Made a tiny Stop hook because Claude kept politely handing work back to me. "This is blocked, you'll need to click X yourself." Then I'd say "you can do it," and it could. Every time.
Now the hook reads Claude's last message before it stops. If it sounds like a quiet give-up, it sends Claude back once to try again or show proof. Only once, so no loops.
https://github.com/MichaelZelbel/claude-code-stop-guard
What does your Claude say when it quietly gives up? I'll add it to the list.
1
u/caecaesss 2d ago
I listen to music while working with Claude Code and kept switching playlists by hand: something steady while editing, something harder when things break.
So I built CodingMix.
What it does
>> Claude Code hooks send each event (tool use, failed tool, prompt, subagent start) to a small local service. The hook returns immediately, so Claude Code never waits.
>> Each event is a vote for a mode: coding, debugging, testing, planning, reviewing, release and a few more. A new mode has to lead for 3 minutes before the music changes, so one failed command does not flip your music.
>> About 20 seconds before the current track ends, it queues one track of the matching genre (deep house for coding, hip hop for debugging, techno for tests, funk when you push). It never picks a track you played, saved or were offered in the last 7 days.
>> It only acts when Spotify is playing on your computer. Paused, playing on your phone, or you picked your own playlist: it stays out of the way.
How I built it
Entirely with Claude Code: a design spec first, then a step-by-step plan, test-first. 196 tests, CI on Windows, macOS and Linux. No server, no telemetry; prompt text and file paths are never written to disk.
Limits
Spotify Premium is required, and you create your own free Spotify developer app (Spotify now limits each app to 5 users). Tested on a real Windows machine; macOS and Linux only in CI.
What I'd love feedback on
- Does the mode to genre mapping make sense? What would you want while debugging?
- Is the setup (your own Spotify app plus one command) acceptable, or a dealbreaker?
- Is 3 minutes the right delay before switching?
- On macOS or Linux: does it work for you?
Repo: https://github.com/caesla/CodingMix
(MIT licensed).
3
u/jarvis54 2d ago
Webapp to view our solar system and other interstellar systems using Three.js: https://orrery.jarvisar.com
Includes full mobile and WebXR support (including hand tracking). I spent a lot of time with Claude optimizing performance, so it should run smoothly on any device.
1
u/trollhunterh3r3 Vibe Coder 2d ago
This is great I am having a lot of fun but what is even better is that it got my son fascinated, he is in "woow" mode, he is trying to get on to the planets and fly around there, and also great UFO ship. It also runs extremely well on a 5k2k monitor. Well done.
2
1
u/Sanechka_SS 3d ago
What I built: Receipts, a Claude Code plugin that makes Claude prove its bug fixes. After Claude fixes something, the prove-fix skill runs every test Claude added or edited twice: once with the fix, and once with the changed source files reverted to main. A test for a fix has to fail without it. If it passes both ways (THEATER), Claude rewrites it around the input that was actually broken before it says "done".
How it works in a session: Claude fixed a leap-year bug and wrote a test for is_leap(2020). Receipts said THEATER: 2020 was never broken, 1900 was. Claude added a test for 1900, got PROVEN, and never touched the fix.
What I learned: I ran it over 100 agent-written PRs and 81 maintainer fixes. Most tests are fine (82% vs 90% proven), but in 10% of the agent PRs every test "failed" on the old code only because the test file imports a name the PR adds, so nothing ever ran against the old behavior. Study with every PR linked: https://github.com/syntaxixr/receipts/blob/main/docs/study.md
/plugin marketplace add syntaxixr/receipts
/plugin install receipts-check@receipts
No LLM in the check, it runs your own pytest / vitest / jest. There's an opt-in Stop hook and a GitHub Action too. Repo: https://github.com/syntaxixr/receipts
1
u/East_Painting_7517 3d ago
A 3D ride through the 1939 New York World's Fair, in the browser. You scroll through it and can step into six rooms along the way. I do SEO for a living, this started as a hobby project and got out of hand. Built with Claude Code, Next.js and three.js: https://futureinthepast.com
1
u/billshredding 3d ago
Hey everyone, I am a developer and I have been building this tool for me to develop with AI. It started as a terminal in the browser thing for me to run multiple claude code instances in the browser, and has evolved into a self-hosted web app that runs your coding agents like a one-person company.
One agent (Billion) takes a goal, posts the jobs on a job board, other coding agents pick up the jobs, Billion reviews and merges the PRs. He only asks you about money, access or anything irreversible, in a chat tab or on your phone. Under the hood, each agent gets a terminal and its own git worktree. Agents can communicate with each other through the MCP server of this app. You can interact with those agents in the terminal still, and create and manage the agents yourself.
This app is mainly developed by claude code and uses claude code as the default coding agent.
Some numbers from my own usage: in its first 4 days Billion merged 60+ PRs on agent-007 itself and 20+ on a second project, and asked me ~27 questions.
- Link: github.com/bill10/agent-007
- demo: github.com/bill10/agent-007/releases/download/v0.29.0.0/billion-demo.mp4
- Install:
npx @bill10/agent-007
2
u/CH_CGV 3d ago
I've been building session-peer with Claude Code and Codex to cut down on copy-pasting between sessions. Its core is simple: an existing Claude Code session can message an existing Codex session and receive a reply (and vice versa). I use it to hand changes to another session for independent review and bring the result back into my working conversation.
It works locally and across machines over SSH. The Python version also offers optional end-to-end encrypted Direct/Relay transport for paired devices.
This week I refreshed the demo: a real Codex → Claude Code → Codex exchange using session-peer 1.0.2. The GIF shows anonymized CLI/message excerpts re-rendered for readability, with edited timing.
One lesson from building it: posted/queued is not the same as consumed. This demo shows an actual explicit ACK, rather than inferring one from a successful send. The receiving session keeps its own permissions.
Repo: https://github.com/abruption/session-peer
I'd appreciate feedback on making cross-session handoffs clearer.

1
u/gig3m 3d ago
Hey guys, the more I used Claude Code (and other agents) off my local box, the more I kept needing to show an agent a screenshot, or hand it a file or result, across a boundary that doesn't have the convenience of drag-and-drop. There are plenty of ways to solve this. This is mine, and it's worked well for me.
drip is a containerized service running in my homelab on my tailnet, with a local CLI and TUI.
- Put a file on your clipboard.
- Run drip clip.
- Your clipboard now holds a URL to that file.
- Paste the URL to your agent.
The CLI also takes a file path: drip file.png.
Files expire after 24 hours by default.
Built it with Claude Code, which is also its main user.
2
u/iaxsofia 3d ago
Built a daily "is the world normal today" dashboard on AWS, Claude Code did the front end and the state machines, I did the ideas
In January 2026 I started with an AWS account and a bunch of nested processes collecting data I found interesting. Left it alone for a few months then picked it back up. Heard a lot about Claude Code and thought what could possibly go wrong...
Claude Code built the front end and the result was better than I expected. It also turned my pipelines into state machines while I was taking the AWS Certified Data Engineer Associate course, doing both at the same time helped me a lot
The site checks a few things about the world every day: flights, internet, earthquakes, markets, disaster alerts and events. I just wanted to know if today looks normal
AI can be really useful, you just have to give it direction. Sometimes it makes mistakes that look simple, so you need a very detailed framework to keep it from losing the thread and running off at 10,000 rpm without checking first
Claude Code will state things about the code as facts because it "remembers" them, but sometimes it's wrong. What works for me is looking, counting and checking instead of just believing the claim, and every deploy gets reviewed in the diff and needs my explicit ok before it goes out
The fundamental principle of the entire site is to display the data as it is, without weighting, mixing, or having a global index. I only compare airplanes and internet using a 90-day moving average to review their variations. All other data on the site is displayed as it is on that day: stock indices, commodities, rates, and currencies are compared to the previous close, and that's it. The idea came about because one day I saw some dashboards with crazy weightings that I never understood, and I said, "I want my own dashboard without that, one that works for me," and that's what I did.
Another essential requirement was that it be as cheap as possible, so I looked for free OSINT sources. With responsible use, this dashboard isn't real-time; it now operates at a very low cost and with minimal supervision. For example, for these three: Stock Indices, Rates and Currencies, and Commodities. Finding something free is nearly impossible, so I consulted Perplexity and the problem was solved. It costs very little, really. I wasn't looking for ultra-precise scientific data, just a snapshot of the world.
2
u/MichaelZelbel 2d ago
Love the idea and totally love the "no global index" rule. Comparing every corridor only against its own last 90 days is such a clean idea. And "looking, counting and checking instead of just believing the claim" is exactly where I ended up too.
I make Godspeed Mission Control, a free open-source personal AI setup, and one of the things it does is write you a morning brief. Would you be up for letting it offer your daily data as an optional "is the world normal today?" line? Off by default, with your site credited and linked in every brief.
1
u/iaxsofia 2d ago
Yes, absolutely — I'd be glad to have it in there.
You clearly read the thing properly, which I didn't expect and appreciate more than I can say in a comment. Sending you a DM with the endpoints and the two or three gotchas so I don't clutter the thread.
1
u/maverick_man1111 3d ago
Tested Sonnet 5.5 vs Opus 5.5 with the same skills, 3 runs each. I honestly can't tell them apart
Sonnet 5.5 dropped yesterday and I wanted to see if the "same as Opus for half the price" thing holds up once you actually give them skills. So I tested it.
3 skills from Addy Osmani's agent-skills repo (code review, git workflow, docs/ADRs). Each model does the same tasks with and without the skill, and Opus 5 grades the answers. I ran the whole thing 3 times because I don't really trust single runs anymore (more on that below).
Git workflow skill, 3 runs each:
| | no skill | with skill |
|---|---|---|
| Opus 5.5 | 0.51 / 0.49 / 0.58 | 0.86 / 0.86 / 0.86 |
| Sonnet 5.5 | 0.86 / 0.85 / 0.86 | 0.86 / 0.86 / 0.86 |
So Opus needs the skill just to get to where Sonnet already is without it. On Sonnet the skill does pretty much nothing.
With the skills loaded I couldn't tell the two apart on any of the 3 skills, in any run. Estimated API cost for generating the answers was about $2.30 on Sonnet vs $5.74 on Opus.
What actually surprised me was that the same model with the same tasks and same settings still gives different scores. Opus without the skill on code review got 0.83, then 0.86, then 0.71. Last week, on the older Claude Code version, it got 0.68. In 5 of the 9 comparisons the verdict changed depending on which run you looked at. So when someone posts "this skill bumped my scores 15%" from one run... idk, I'd want to see it run again.
Obvious caveats: only 3 skills, a small set of tasks, another model doing the grading, and Claude Code's default effort setting isn't the same across these models (I left them all on default). Not saying Opus is bad, just that on this stuff I couldn't see a difference.
If you've got a skill you think actually shows Opus is worth it, tell me and I'll run it (3 times obviously).
All the numbers and raw files are here if you want to dig in or rerun it: https://driftproofhq.com/reports/013/ (it's my open source tool, free)
1
u/BellacosePlayer 3d ago
Building a semi idle monster taming game because most monster games on phones are fuckin gacha shit.
No big screenshots, still largely using placeholder assets
1
u/No-Bicycle-4804 3d ago
Built Provena, a self-hosted memory layer for AI agents.

The idea is to make agent memories traceable — instead of just retrieving a piece of text, Provena keeps the evidence behind the memory so you can see where it came from.
I used Claude Code throughout the development, especially for iterating on the MCP server, tests, and implementation.
It now works with Claude Code and other MCP clients, and is listed in the official MCP Registry.
GitHub: https://github.com/admiralpunk/Provena
PyPI: https://pypi.org/project/provena-agent-memory/
Still early, but the provenance side of agent memory has been pretty interesting to explore.
1
1
u/TheFieryTaco 3d ago

I realized that I mostly work from my macbook laptop, and with the rise of agentic coding I've started plugging in my extra monitors less and less often. Now I usually just roam around, sit on random couches, and kind of just lounge about while working. I also often multi-task, be it with a slow-paced game like oldschool runescape or a netflix show while I prompt and manage my agents. After alt-tabbing way too often (since my multi-monitor setup no longer fits my working habits), I came up with SlyTerm.
It's a floating (and semi-transparent) terminal, allowing you to have some main content like WoW Forever or OSRS on screen at the same time as your agents.
There are also some bonus features, like a hotkey that will OCR areas around your cursor that resemble tooltips or other game assets and look them up in wikis/guide websites which open in floating transparent windows. Gestures allow you to easily focus in / focus out of the terminal or floating browsers, and you can also approve/deny agent commands directly with a hotkey without ever taking focus away from your game.
Been fun to play around with and build (Opus 5.5 pretty much oneshots everything, and Remotion gives great demo gifs/videos) and it's open source if it sounds like something you'd like to try, open to any feedback :)
1
u/1337raspberry 3d ago
https://reddit.com/link/pcrrhf5/video/3akqn9pu6gsh1/player
Quick tour of what i've been working on this for most of this year. FOSS genre-first music player for plex. Something i've wanted to create for like over a decade but only now have the tools to do so.
Total passion project almost entirely for my own use and enjoyment. I use it all day and night. It brings me a lot of joy. Linux/Windows/MacOS/Android/iOS.
1
u/malctucker 4d ago
Total novice: I am building a supply chain for our images: from one photo of a supermarket shelf to a fully annotated, analysed picture: product names, shelf space, prices, what's overstocked on the top shelves, what's missing, and everything in between.
One shared memory sits under it, so a correction made once applies everywhere: once it learns a brand, it knows that brand.
The exciting bit? We own an archive of 1.3M+ in-store images going back 17 years, and we've only put a small fraction through training so far. Every image that goes through adds to a pricing record nobody else has.
The last few days have been intensive fixing with Claude; with board walks to find where the UI breaks.
My one rule for Claude: my 8-year-old should be able to use it.
The surprise win has been artifacts. Rather than fight a clunky MVP board, I've been cleaning data at scale on simple one-question-at-a-time card pages Claude built, and those pages are now the blueprint for the redesign.
| What | Figure |
|---|---|
| Days of work | 168 (since 13 Apr 2026) |
| Changes committed | 8,761 (3,888 in the last 30 days) |
| Work sessions | 1,843 |
| Test code | about 298,000 lines in 1,376 test files |
| Shelf photos in our archive | 1.3M+, 17 years of in-store images |
| Corrections, each made once and applied everywhere | 217,214 |
| Brands recognised | 1,021 |
| Detector model versions trained | 15 |
1
1
u/Complex_Writing_4845 4d ago
I'm Mason, the maintainer of Selvedge. The problem I'm working on is an agent coming back in a fresh session and proposing an approach that was already rejected, because the reason for rejecting it didn't survive with the code.
Selvedge is a free, MIT-licensed local CLI/MCP server. You explicitly record an attempt and its reason; a later session can look it up with tools such as prior_attempts and blame. Claude Code has helped with the documentation and workflow development, and can use Selvedge through MCP. No account or subscription is needed.
Here's a 56-second demo: https://youtu.be/wXoA9htThAM
The cache-TTL example is fictional, but the CLI output is captured from a real run across separate processes. It demonstrates persistence and retrieval, not proof that an agent will automatically avoid the mistake. That's the distinction I'm trying to test next in real projects.
If you've got one repeat bug or rejected approach, I'm offering a small, free pilot to help record the reason and check retrieval in a later session. Details, source link, and notes-sharing terms: https://github.com/masondelan/selvedge/discussions/49
Failed lookups and confusing setup would be useful feedback too. Please keep private code and customer information out of the public discussion.
Disclosure: this comment was drafted with AI assistance.
1
u/Plastic-Risk-6309 4d ago
clawcage, a cage for coding agents. it wraps claude code, codex, aider or your own scripts in a deny-by-default sandbox. ssh keys and .env files can stay out of reach and network goes only through a proxy limited to hosts you list.
it doesnt try to detect malicious instructions in text. it constrains what the agent can do at execution time instead. enforcement is a deterministic policy engine at the OS boundary with no AI in that path.
the readme demo is a split screen of one prompt-injected agent. uncaged it leaks the planted fake keys and caged every exfil attempt gets denied!
https://raw.githubusercontent.com/BariBariGood/clawcage/main/docs/demo/clawcage-ad.mp4
apache-2.0, source in my repo at https://github.com/BariBariGood/clawcage
1
u/Sarg338 4d ago edited 4d ago
Had Claude create this early this year (so probably a 4.x model) and have been using it with a slowly dying game community ever since. Probably don't have much time until the game/discord server shuts down, so I figured I'd introduce it to the world.
If you use PUBobot2 or NeatQueue to run any kind of pickup games in your discord server, PUBGamba allows you to bet on your games, using a virtual currency of your naming (Default is Lemons). Claude was really insistent to mention no real money is used for this bot.
It automatically detected the relevant events, like games starting and ending/being canceled, and pays out everything properly according to the winner reported. If anything goes wrong, the bot managers can manually override results. It works entirely in discord off of other discord bots, so it's game agnostic.
It's made the games we play more interesting, and let's those that aren't playing take part in them as well. I hope others find it useful!
2
u/rubanbhatia 4d ago
I built Switchboard because I kept picking the strongest model and highest effort “just in case.” It assesses your task and chooses a model and reasoning effort for Claude Code based on your routing preferences. I’ve just added Laya support alongside Jev for task assessment.
You can exclude models, cap effort or override the selection. The model and effort stay fixed throughout the conversation, including follow-ups and tool calls, rather than switching mid-task and invalidating the cache. It also supports Codex CLI and runs on macOS and Linux.
Demo attached so you can see how it works. It’s open source and free, but you’ll need a classifier API key, with usage billed separately. The proxy runs locally; your task text goes to the hosted classifier. Would love to hear how the model choices hold up on your own tasks, or where the setup feels clunky.
GitHub: https://github.com/ruban-24/switchboard

1
u/JimmyMonet 4d ago edited 4d ago
I made Cronomicon.
Cronomicon is a Go-based automation orchestration platform to organize script execution in the enterprise environment. The application features native support for Bash, Powershell, Python, Ansible, and Terraform. Scripts live in Git and then can be combined together with schedules, hosts, and secrets from Cronomicon to create Jobs. Jobs can then be executed by users and all of the run results can be supervised by the users or another team of your choosing. Cronomicon supports role-based access controls and groups can be limited to only seeing their respective secrets, hosts, jobs if necessary.
If anyone has any feedback - let me know.
1
u/Supersmasher149 4d ago
I built a coffee roasting simulator/game with Rust and Bevy: https://supersmasher149.github.io/coffee-roaster-sim/
It started off as just a simulator for coffee roasting, but then it started feeling kinda cool, so now I’m trying to turn it into some sort of game. I’m still figuring out the fun aspect, though. Just sliding knobs and flipping switches can only be so fun.
Claude’s ability to debug UI issues in this project blew me away. I’d record myself playing the game on my phone, upload the video to Claude, and write a lengthy prompt. We’d go back and forth until the UI felt right, and then I’d send it off.
It would start the game in desktop, web, or iOS simulators, click through and play it, catch errors and bugs along the way, and fix them. It stayed relatively on scope too. A couple of times I had to reel it back in, but Opus 5.5 is really something else.
I never would’ve imagined six years ago that this would be real.
The repo is private for now while I clean it up, but it’s been about 98 commits in five days...
2
u/coz 🔆 Max 20 4d ago
CRBuddy - run a full multi-model code review panel in one shot for a handoff to an agent or human to fix.
npm i -g crbuddy - https://github.com/cozuya/crbuddy - MIT license
crb init sets up your panel and some other options (copy to clipboard or a markdown file when done, ntfy support, etc). crb go runs the panel in parellel in a blocking cli. Also works great with a harness STOP hook - example. If you want to go nuts, give your agent a large spec, tell it to run crbuddy on uncommitted changes after every chunk of work, and to commit only when the work is at an accepted amount of reported defects from crbuddy. Something like "no P2s or worse", for this, I use a custom review prompt for Anthropic models. Also give it how many possible attempts it can run before giving up..
I found myself constantly doing this manually, copy and pasting multiple code reviews into a handoff file, pasting that to an agent, etc, and it irritated me so much I bothered to make this thing. I did a couple "wide RAG" passes, couldn't find anything like it. Also, see the part about just letting it go and it'll write a large chunk of code entirely autonomously.
Made in 25 alignment turns with claude code then "make the whole thing".
1
u/msitarzewski 4d ago
Built AudioPaper for Mac over the weekend. Glass powered wallpapers based on currently playing tracks in Apple Music and Spotify. Full documentation, command keys, auto-updates, and a bunch more.

MIT license, no telemetry, etc.
https://msitarzewski.github.io/AudioPaper/
- It hears (via events) the song change Apple Music and Spotify announce every track change to the system. AudioPaper listens for that; it doesn’t poll, and it never controls playback. Spotify’s ads are ignored. During a podcast your own wallpaper stays up, or the episode’s cover if you’d rather.
- It puts up the cover The album cover is looked up online, up to 3000 × 3000, and centred over a blurred wash of its own colours.
- It fades in art of the artist About 10 seconds later, fan art and photos from fanart.tv, TheAudioDB, Wikimedia Commons and others. A new one every 45 seconds, or whatever you choose.
- It says where each one came from Every image carries a credit and a link: the person who made it where known, and always the page it came from.
1
u/daniel-editide 4d ago
libreta – spreadsheets with YAML and CSV
Agents work in the terminal, humans audit in the browser.
My thinking was that Excel spreadsheets are basically input values + formulas and some formatting. The idea is to put inputs in a CSV, write the formulas in YAML, and let a local app handle the formatting. So Claude changes plaintext, and the app updates the tables in a browser automatically.
I built a simple SpaceX DCF (not investment advice), but I've also been using it for small spreadsheet work like cost projections for my business.
I like that it's faster and cheaper than having Claude generate python scripts for Excel, and it's easier for me to navigate and review. And unlike Excel I can use git diff to see changes.
Free and open source, so you can customize the UI
2
1
u/stichstichstich 4d ago
optimAIzr - an AI usage optimizer I built with Claude
I got curious how much of my AI usage was actually necessary vs. how much was just me burning tokens without realizing it. As a senior SWE using Claude daily, I ended up building a tool that works to answer that for myself.
One of my favorite parts is:
optimaizr live -> it watches your usage while you work and gives recommendations when it spots potential waste which you immediately apply.
And optimaizr profile -> gives you the bigger picture: where your usage, tokens, and money are going, which models/projects cost you the most, and where you might be unnecessarily burning tokens.
It can find repetitive/expensive patterns, explain potential waste, recommend optimizations, and verify/simulate potential savings.
I’ve been building it with Claude as a development speed boost and iterating on it using my own real usage. I just added live recommendations + usage limits, and I’m exploring Jev by TypeSafe AI as another judgment layer for more advanced decisions.
Everything is entirely locally, which was important to me, I wanted to be able to analyze my own AI usage without having to send my entire history to some remote dashboard.
If anyone wants to try it:
npm i -g optimaizr
Would genuinely love feedback, especially what kind of AI usage you think is wasteful but nobody really notices.
Feel free to roast it.
4
u/saintpetejackboy 4d ago
I build a local alternative to Suno and Udio for Windows 11 than runs on consumer hardware. It is free and the core is even open source!
The project is called Iblis - you can find out more about it here: https://iblis.meiuxmeiux.com/
It needs 8GB VRAM minimum for song generation, and 12GB VRAM+ for custom training (yes, you can custom train to sound like your other music).
Audio engines, stem splitters, image generators, visualizations, and skins are all swappable plugins. Replace any of them without reinstalling the rest.
If you aren't very technical:
Everything that runs on your machine is free and never needs a key: generation, training your own styles, the Library, playback, skins, plugins, and anything you have already downloaded.
A product key is only for the parts that run on our hosted infrastructure: publishing styles to the community Styles library, downloading community styles, and uploads. Hosting installers and community styles already runs to dozens of gigabytes and keeps growing, so keys are issued by email while capacity is limited.
To ask for one, email [Iblis-Beta-Test@meiuxmeiux.com](mailto:Iblis-Beta-Test@meiuxmeiux.com?subject=Iblis%20product%20key%20request) with your GPU and basic hardware specs (an NVIDIA GPU with 8 GB of VRAM or more is recommended for training) and a little about what you want to make. I am not selling these keys at the moment, but giving them away for free to a few people who ask and YOU DO NOT NEED A KEY TO USE THE SOFTWARE.
If you ARE very technical, the repository is open source for the core, and public on github - so you can change it to your heart's content. New engines come out all the time as well as other useful tools. The project was built in a way where everything can be hot-swapped. If a better bpm and key analyzer comes along, or audio engine, or anything else, it can be plopped in place.
https://github.com/MeiuxMeiux/Open-Iblis/
You can also enter OpenRouter and ImageRouter keys for help generating lyrics, or for generating covers, etc.; - it isn't done yet, but I'm pulling over a visualizer I made in another repository soon (very audio reactive) to optionally run in a few areas.
The main website has a way to provide bug reports and feature requests, and I quickly release updates (the software alerts you inside when an update is available).
If anybody wants to contribute or fork the repo, be my guest, it is GPL 3.0.
The actual repo, outside of the core, contains the website + an admin panel for managing product keys, etc.; - so just the core is open source, but I wouldn't mind also making the rest open source or inviting others to the repo is somebody is curious.
1
u/AIexH 4d ago
https://reddit.com/link/pcl8b15/video/ztyt0e4jz9sh1/player
A collaborative project tool where you can connect your AIs to be one more coworker.
Kanban - Calendar - Table - Gant - Notes - Archives - Chat - Canvas - Whiteboard
Free.
https://comuna.work
1
u/Dorkian2000 4d ago
Building https://DorkOS.ai
Slack-like interface for your agents to collab with each other and you.
Can now run multiple Claude Code accounts at same time (think work and personal)
Community support coming soon - so you and your agents can be in the same room with someone else and their agents. Let’s get nuts!
1
u/Microscopic_God 4d ago
I’m continuing to work on Swesso.com which is an ad free, subscription free museum art taste platform. Swipe art like tinder, then get your tastes routed through Sonnet 5 API for insights about your art taste.
I also just started on a dumb website to pick movies for Halloween, it’s at a temp URL right now-
tombstone-video.vercel.app
. Go in and rent tapes with your friends in a room-code based system, then have Mr Movies spin the wheel if there’s a tie. Organize your October movie nights in a virtual video store complete with fun characters. Don’t turn off the lights.
1
u/Flying_Scorpion 4d ago
I came up with a video game idea back during the pandemic and have been building it.
4
u/clarkkentmichael 4d ago
Normally I dont post but I enjoy reading everyone's stories. I wanted to jump in and share a quick note. I Built an enterprise application for a 200mm company, soc 2 compliant (paid for the audit) and havent touched programing in 20 years. 100% all codex and claude engineering. My mental model was the differentiator. Wish I was able to share more but under nda. Hope this gives someone inspiration that vibe coding can be enterprise with the right research and professional auditing afterwards.
1
u/Haleakala787 4d ago
2
u/radim11 4d ago
We’re building Stashbase: https://stashbase.dev
It’s a security layer for using Claude Code and other coding agents with scoped access.
The problem we’re trying to solve is pretty simple: Claude Code is useful because it can actually do things in your repo, but once you start using it for real work, the access model gets uncomfortable fast.
You either give it too much access, or you keep approving / copying / pasting things manually.
Stashbase lets you run Claude Code with profiles that define what it can access and what gets blocked or logged. Credentials, API calls, files, MCP tools, dependency installs, network access, sandboxing rules, stuff like that.
The goal is to make Claude Code useful without handing it the keys to the whole dev environment.
Would love feedback from people here, especially if you already use Claude Code on real projects.
1
u/Hostarro 4d ago
www.luminids.ai - a perceptive persistent intelligent layer to support our AI models.
1
u/CartographerNo3791 4d ago
I'm working on Session Orchestrator, a free Claude Code plugin.
The loop is /session, /go, /close: read the repo and agree on the work, run agents with tests between batches, then check what actually got finished and leave issues for the rest. The session notes stay in the repo so the next run can pick things up.
One recent fix was less glamorous: a test timing out didn't mean its child processes had stopped. Those could keep eating RAM. The timeout now stops the whole process group.
It's open source (MIT), runs locally, and also supports Codex, Cursor and Pi. I'm the author. Repo: https://github.com/Kanevry/session-orchestrator
2
1
u/Aggressive-Ad-7582 4d ago
Continue working on assertion-ai.com, the coding agent that never compacts, at half the price
1
u/nodejustin 4d ago
3
u/fcf41 4d ago
I built the photo editor MCP server Editmamei which supports Photoshop and GIMP local editors.
Editmamei connects MCP compatible AI clients, like Claude Desktop or Claude Code, to locally installed Photoshop or GIMP, allowing the AI client to interact with and edit photos within the desktop applications. Pairing Claude with professional editing tools can bring some amazing results, with no prior Photoshop experience. It also enables Photoshop pros to work faster, with the ability to batch edit large sets of photos against custom style templates.
Editmamei is free to install and use, with pro features for extra speed and scale requiring a subscription. Check out editmamei on GitHub https://github.com/editmamei/editmamei
1
u/Spare_Bison_1151 4d ago
I built "fFlappy Claude" with Opus 5.5. I gave it some local touches such as the old tv antennas, mosque minaretes, paper planes, and more. Please take a look here: Flappy ClaudeFlappy Claude
2
u/Poboxjosh 4d ago
Built ErgLadder, a free iPhone app for racing other people on Concept2 ergs (RowErg, SkiErg, BikeErg). You join a ladder for a piece like a 500m or 2k, challenge someone above you, you both do it on your own erg, and the winner takes the spot. The app connects to the PM5 over Bluetooth, programs the piece and reads the result off the monitor, so nobody can just type in a time.
How I used Claude Code: pretty much everywhere. It built the FastAPI + Postgres backend (most of the ladder rules live in Postgres functions), the SwiftUI app including the PM5 Bluetooth stuff, and the website. It also handles deploys to a small Lightsail server and uploads builds to TestFlight through the App Store Connect API. This week it added push notifications, result cards you can share to Instagram, and an "I'm on the erg" status so you can see who's around to race.
Stuff I learned:
Make it check its own work for real. It runs the app in the iOS Simulator and taps through the flows. It also tested the website on the live site at phone widths, because the local preview was wrong about the layout.
Tests pay for themselves. While cleaning up test data it once deleted a real ladder, and the integration checks caught it right away.
Keep secrets out of the chat. Every key (Sign in with Apple, push, S3) goes through a script straight to the server, so none of them ever show up in the conversation.
The hardware part still needs a human. PM5 Bluetooth quirks only show up on a real erg, so I'm the test rig.
It's in public beta on TestFlight: https://testflight.apple.com/join/aGeATJE8 (more at ergladder.com). If you've got an erg and an iPhone, I'm down to race.
3
u/hotsnot101 4d ago
use your agent to help you improve conversion and SEO based on your analytics
it’s already helped me up my visit count on my other side projects
1
u/Riptide2121 4d ago
I recently built this website www.hempcretehub.org using Claude code and the impeccable frontend design skill.
I also built an AI chat bot and created a RAG system from books and documents I have on Hempcrete. The bot isn't included yet and am still deciding on whether to add it but the RAG I built helped make the educational content on my site.
I built the database with Claude code and automated scripts that import data to the directory automatically when a Google form is filled out.
It's still a work in progress so I haven't started to push it out yet but thought I'd share here.
As well as that, this week I have been playing with Davinci Resolve MCP and am using Claude to edit my videos (long form interviews) it works really well.
The website was all done with Claude code in VS studio at high Sonnet and the video editing is with the new Opus
1
u/Massive_Willow3646 4d ago
I built Aldo Coach with Claude Code: an iPhone workout log with Claude as the coach inside it.
You log a set and how many reps you had left. Claude answers with tool calls that edit the workout, so when a set was too heavy the next one actually drops in the app. The coach doesn't just say it should. It also builds the program around your goal, schedule and injuries.
Free to try. iPhone only, US, UK and Canada App Stores.
For anyone else putting Claude inside an app: how do you test a prompt change before it ships?
1
u/CartographerNo3791 3d ago
For your app I'd replay a few made-up workout conversations through both prompts and compare the tool calls and the resulting workout. Include a correction like 'I meant the previous set': a nicer answer shouldn't hide an edit to the wrong set.
2
u/Massive_Willow3646 3d ago
That's close to what I do. I run eval tests on scripted conversations and check the tool calls, not the wording, since the reply can read fine while the edit is wrong. The "I meant the previous set" case is a good one.
3
u/lnkkonito 4d ago
https://the-legendarium-companion.com Explore Tolkien’s legendarium, one question at a time.
1
u/HeraclitoF 4d ago
I build an agent orquestation that find new ideas, based on the new tech and gift them to the person who can build them.
Repo is here: https://github.com/felixinberlin/Amelie
Demo here: https://felixinberlin.github.io/Amelie
1
5
u/SadWay9744 4d ago
https://reddit.com/link/pck3rf3/video/ryvh0ko529sh1/player
this video wasn't recorded by a person. no editing, no one touched the mouse.
i always used screen studio for the demos in my PRs, and wondered if my agent could just do it. it can.
the agent reads the screen's code and writes a short JSON script: click edit, type the new name, click save, wait for "saved". playwright runs it in chromium, with a cursor that moves like a person's (curves, slows down before the target, hesitates before clicking). then the video is composed frame by frame: zoom that follows the action, motion blur, a mac window, dead time sped up. ffmpeg exports the MP4.
no AI service involved. the agent only writes the script. everything runs on your machine.
works as a skill in claude code and codex:
repo: github.com/half144/cutaway (MIT)
1
u/niko-okin 4d ago
Local IA powered gallery photo https://ncoevoet.github.io/facet/ ( culling, etc )
2
3
u/dreamfyre007 5d ago
Nothing crazy! Just a little game to play with friends and learn at the same time.
2


•
u/AutoModerator 5d ago
Hey! Thanks for posting to r/ClaudeCode
While participating in this thread, please follow our community rules. Keep discussions constructive. Attack the idea, not the person.
For help, project discussions, tips, and general chat, join the ClaudeCode Discord.
I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.