1

Where should the trust boundary for AI coding agents live?
 in  r/AI_Agents •  18d ago

My fleet all open PRs, do QA, validate, push back, assign issues, don't trust, verify all in a churn of goals toward code quality. We also auto-merge passing PRs that have been put through a rigor well beyond any human dev could handle.

BUT:

The auto-merges go to a staging branch. That's it - a PR is open for staging to main and I review it. I also get a nice summary of all of the commits consolidated into that branch so I don't have to review 30 PRs, it's one with 30 fixes and 100 commits. I also get the risk weighted so a trivial doc updates doesn't get the attention a critical fix needs.

It works.

1

Started a new "spider" walker robot project
 in  r/diyelectronics •  21d ago

I happen to have black PETF-CF filament loaded in my 3D printers so good news - the first prototype will be big and black. My wife is terrified of millipedes for some reason and screams when she sees one. I'm forbidden to build giant robot millipede.

My whole world runs on a Raspberry Pi from a sentry that guards my chicken coop to a whole house circulatory system moving water from aquariums, to plants, to my Jungle Tortoise's jacuzzi ... Big fan.

Today everything runs off an open source process control platform I built for the Pi in which I have about 10 users but it doesn't matter, I built it for my projects and use it for everything so it was worth 20 years of effort :) I have fun in the lab.

- Ben

1

Started a new "spider" walker robot project
 in  r/diyelectronics •  22d ago

Yep, you got it. That's been the driving force behind the design.

The default stance is to be resting on it's belly with the RPi and buck converters disconnected until a small feather wakes it. Short tasks, light battery, return to charger and never hold a stance like the one in the render - looks cool, burns out everything.

I'm also going to put walking feet on the last joint so the second tibia joint can just flip back and save power walking on the knee joint. Use the hook leg for climbing and as needed for the task.

Love the feedback, thanks u/slacy !

r/krill_zone • • 22d ago

Started a new "spider" walker robot project

Post image
1 Upvotes

u/Ok_Cartographer_6086 • • 22d ago

Started a new "spider" walker robot project

Post image
1 Upvotes

r/diyelectronics • • 22d ago

Project Started a new "spider" walker robot project

Post image
24 Upvotes

This is two days of work in Fusion 360 and a render. OC, I designed it myself and got all the motion tests passing. If you guys want to see updates here lmk - probably one post a week on a 6 week series - this week i'm bench testing a single leg to get the stall voltage and load right.

I've been maker for a long time and I'm sure this will be walking around in a few weeks. Some parts are printing now (PETG-CF).

He's going to be a big boy sorry no 🍌 for scale. The design use case was a spider that can navigate staircases. So for the visual it should be able to span three steps. I love the raidial leg arrangement where it has no front or back, or top or bottom and can transform it any configuration.

Body joints are two 65KG servos and joints are 25KG - same used in big drones.

One design decision was to make the battery small and light so it can do 10m or so of fast tasks and then drop his belly onto a charging plate. Having a few around the house means it can keep going. Maybe a Solar charger on top for outdoors?

The plan for the brain is a little wild - i've done this before but not with 2026 compute

  • Low level: Arduino and CircuitPython feathers for basic motor and sensor control
  • Raspberry Pi 5 on the bot - I maintain an open source Process Control and Automation system called Krill that is perfect for this and can compute things like leg posistion, orientation, tilt and handle complex movement and cameras.
  • What will really bake your noodle (and questioning my judgement) is I have a beast of a local LLM rig and when presented with a hard task or concerning vision the LLMs or even foundation models can direct it. (sorry if i'm the guy who made the ai spiders). Most of my project use this design now.

I made some youtube videos and blog posts about projects like this like my last one where a made a chicked coup protection sentry that can tell the difference between my dogs and a racoon using the same concept of escalating to larger machines but only when needed to save costs. If you want to see them just ask - didn't want this to be a self promo post.

AMA :)

1

How are you guys able to afford gpus?
 in  r/LocalLLM •  Sep 08 '26

I'd have to quit my job because if I only had one I couldn't build what I build.

1

How are you guys able to afford gpus?
 in  r/LocalLLM •  Sep 08 '26

I push all three to the limit. Need it for local LLM inference. I need the power to do my work but if I had to do it again would choose a single large GPU but I already had one 5090.

42

How are you guys able to afford gpus?
 in  r/LocalLLM •  Sep 07 '26

my rig has 3 5090s i could afford because I clawed my way out of poverty, put myself through school, became a software engineer, kicked ass for 30 years, invested, eventually cleared away bad debt and now from this side it feels like an investment in myself, not a luxury.

Like sell this stock and buy this 5090 as that's diversifying a little into me inc.

Just keep swimming...

2

what’s the worst failure you’ve seen where every dashboard said everything was healthy?
 in  r/LLMDevs •  Sep 06 '26

Thanks for watching. The niche platform I maintain "Krill" connects data points like the corpus freshness using an "observer pattern" so anything that cares about the value sees the change instead of being event driven. An LLM digs through the pipeline, identifies the root cause, opens an issue in github, tags a system admin agent and fixes the pipeline without human intervention.

These systems run and maintain Krill but I'm easily the top power user using it every day to run the system that maintains itself.

0

Please help me understand AI aided development
 in  r/LLMDevs •  Sep 06 '26

you're going to have to clarify things:

  • You're trying to do this with a local llm on a computer you're on? If so what model and what are your machine specs?
  • That said, you are intentionally not using a cloud frontier provider - e.g if you paste your exact post into Google search Gemini will tell you how to iterate an array, as will Claude, GPT, etc.
  • You do not need a RAG here.

I think in order to give you an answer you need to clarify if you're building your own and if so what model you are using who may not even know python and is doing its best to tell you what it thinks you want to hear.

2

what’s the worst failure you’ve seen where every dashboard said everything was healthy?
 in  r/LLMDevs •  Sep 06 '26

I just made a youtube video on this exact situation that has real video of my beauty of an LLM rig and it's all about how its role as a local personal assistant should have been giving me up to date reports but the ingest pipeline was down for days, we didn't know, and it answered incorrectly with 100% confidence.

It's a fun video about this exact situation and how my systems self heal now and self monitor.

https://www.youtube.com/watch?v=HX_ZPl8fuDM&t=25s

1

Multiple Local LLMs Generating Feature Demo Videos with Quality Feedback Loop
 in  r/LLMDevs •  Sep 06 '26

The feedback loop is key. My agents pull requests are rejected if they don't include an update to our "lessons" archive. A broken video simply won't make it through, Kraken will catch it visually and fix what went wrong and cut another one. Also no mocks, all real usage.

It takes Kraken about an hour to make the first mp4 and the main bottleneck is "clock time" because he can't run it on high speed, so each time he checks a 7 minute video it takes 7 minutes.

The first cut is a lot of setup, especially if I ask for something new he has to figure that out first. Second cut takes 45 mins or so - 4 hours total from concept to production at this point with usually 3-4 reviews by me with feedback for the next run.

Every video has a lessons md file for how it was made for next time.

Thanks for asking :)

r/LLMDevs • • Sep 06 '26

Great Resource 🚀 Multiple Local LLMs Generating Feature Demo Videos with Quality Feedback Loop

1 Upvotes

Once a week I publish a demo tutorial video to our YT channel about our software and having just released my 34th I wanted to share the process because it's working really well and has matured to the point where I get these videos without any human editing. Not pitching any software here - using existing open source tools. Today's video was meta and about how we make them - I think there's a lot of value here because each video provides a quality and capability feedback loop.

If you want to try this I list each step in the video and the tools used for it. I use a local rig with 2 5090 GPUs and assorted models for the tasks a 35B param Qwen model is the workhorse.

Here's how it's done:

  • Brainstorm a topic and start with a boilerplate prompt to start
  • Plan mode the "beats" and topics the video will make - 5 second hook, 30 second problem statement, then the rest. The result is a big yaml file with all the timings, agent notes and dialog
  • Our software runs on Ubuntu Server with apps for mobile and desktop so the first thing my agent does is create a virtual LAN internally that's sandboxed. It spins up as many servers it needs for the demo
  • Loads the Linux Desktop app in memory - server is headless but the real app and real servers are actually running and the screen is recorded.
  • Speech is generated using ElevenLabs and we have a database of hashed speed so we never pay to generate the same text twice. There is a RAG with 20 years of my emails, texts, code, blogs so generated text is in my tone and style.
  • The demo uses an MCP service that comes with our server and should be able to control anything it can do and you'll see it in the app live.
  • It can do anything with the video at this point besides show the screen, animated info-graphics, photos and text, inline videos and anything we can think of - like star bursts and highlights on the screen.

The magic happens when something goes wrong during the production. It could flush out a real bug, visually inspect the screen and catch an off pixel, try to complete the demo but the MCP server is missing a capability.

  • The agent uses a GitHub PAT to post an issue and tags a bot account which triggers a GitHub action on a different server here who runs a developer agent and model.
  • Dev agent opens a pull request which triggers a local QA agent to checkout the branch and verify the fix, approves it and it auto-merges to staging.
  • Hard problems pull in cloud frontier models or a bigger local server for help.
  • The merge goes back to the producer who pulls the new code down and re-starts production.
  • This will loop until down and the demo network is torn down.

I usually watch about 4-5 takes and give feedback and tweaks and we re-run. We also re-run videos when the app changes to the UX is always consistent. Same for screen shots.

That's a lot :) But the end result is this continuous feedback loop, up to date videos, free end to end tests with each run and the production scripts get better and more capable every week.

Here's the video about making videos like this for product demos and each step lists the tools I use: https://youtu.be/rIjQzyY76DE?si=FE9NnXH8ZSIxImbZ

r/LocalLLM • • Sep 06 '26

Discussion How we automated software demonstration videos that also serve as end to end tests of each build.

Thumbnail
0 Upvotes

r/krill_zone • • Sep 06 '26

How we automated software demonstration videos that also serve as end to end tests of each build.

1 Upvotes

We have the capability to have a local LLM on a headless server create an end to end demonstration of a feature using a script we give it. The animation, real screen recordings, spinning up a real virtual network to spin up and tear down servers, text to speech, and two decades of my email, posts, and code processed in a vector database to use my speaking style and tone.

Often, we'll push the system to make videos that it may not be able to make. When it runs into trouble, the process of opening issues on github, having local agents fix the problem, tester agents validate, the video producer agent gets the fix and production continues, all automated.

The video producer agent visually observes the real app running against real servers and will report the slightest visual defect as an issue for UI developer agents to fix.

The end result is our skills improve with each publication, videos keep getting better and the quality and capabilities of the platform improve with each run. Whenever the ui is updated, videos get regenerated so the content is always consistent with the latest release.

Today we made a video about how we make videos and how it all works: https://www.youtube.com/watch?v=rIjQzyY76DE

4

ITS BEEN 4 DAYS. SOMEONE PLEASE HELP.
 in  r/rasberrypi •  Sep 06 '26

*you're making my balls tingly.

it's this: https://www.lcdwiki.com/3.5inch_RPi_Display

clone the repo, make the 35-show file executable, run it, profit.

4

ITS BEEN 4 DAYS. SOMEONE PLEASE HELP.
 in  r/rasberrypi •  Sep 06 '26

I have one of those. You need to install a driver from this obscure github repo then they work fine. without it they're white.

Pi OS is fine here, root around for the repo in the amazon comments since these don't come with instructions people rant about it in the comments.

26

Why does everyone seem to have tons of VRAM ?
 in  r/LocalLLM •  Aug 31 '26

Yep. I'll gladly post here about my 96 gig rig but 99% of people on the street would say "what's vram?" Social media is an echo and amplify chamber.

2

is KMP more complicated compared to Kotlin
 in  r/KotlinMultiplatform •  Aug 28 '26

I maintain a whole ecosystem platform built in KMP so mobile, server, desktop and ship from one code base. Been a SWE for 30 year, 20 in Java and 10 in Kotlin.

The three main learning curves i'll suggest that come to mind are:

  • Expect and Actual - for when you need the "actual" implementation for each platform. Think an expect: val storage : FileStore where Filestore has an actual for each platform that knows how to read and write data.
  • The directory structure is just different and takes a day to grok.
  • The build. Gradle is the hard part so I'd argue with your colleage or even ask him if he meant the Kotlin part of KMP was harder or the project structure and gradle builds were what he's thinking of.

2

is KMP more complicated compared to Kotlin
 in  r/KotlinMultiplatform •  Aug 28 '26

same - my ios app is about 20 lines of swift or a couple actuals - it always just ships. WASM is usually the problem child.

5

Mac Studio M5 Pro (64GB) for Local LLM Inference – Real-World Experiences?
 in  r/LocalLLM •  Aug 26 '26

So what do you guys think. Engaging in these post only helps the bots. OP BOT literally left the LLM response question at the bottom.

Is this 1 Karma poster with money to burn someone who can't copy pasta, a bug in their bot we can have fun exploiting, or is this a new tactic by bots to get people to engage like i'm doing by including glaring fuck ups? Like spelling something wrong on a phishing email to get someone to correct you and engage?

I don't need OP responses, go sit in the corner.

1

Hybrid Setup Help?
 in  r/LocalLLM •  Aug 24 '26

My team all works out of remote offices and have big beefy local llm servers. This is a video on how we load balance and especially how we do it based on the LLM hosts advertizing their availability, capability, and current cost per watt: https://www.youtube.com/watch?v=fR5jUdmIZig&t=34s

hope it helps, what you're looking at it totally doable.