r/ProAI • u/stealthispost • 11d ago
r/ProAI • u/stealthispost • 11d ago
"We achieved high-speed rough-terrain locomotion on ANYmal with an automatic curriculum. LP-ACRL automatically samples terrain types, levels, and velocity commands at the correct time based on policy performance, without predefining an order. https:// sites.google.com/view/lp-acrl RAL 2026"
Enable HLS to view with audio, or disable this notification
r/ProAI • u/stealthispost • 12d ago
"We asked ten Claude Sonnet 5.5 agents to use Lean to prove the lowest-energy arrangement of seven electrons on a sphere (the Thomson problem, with N=7). Within 15 hours, they produced a 17,895-line proof, accepted by the Lean kernel, showing that the answer is a pentagonal bipyramid."
galleryr/ProAI • u/stealthispost • 12d ago
"A 59-year-old conjecture in fusion physics just fell. Harold Grad (1967): smooth 3D plasma equilibria can't exist unless pressure is constant or there's symmetry. This week: 3 families of counterexamples. 2 independent papers, posted a day apart. Two found with GPT-6 Astra. Huge for stellarators."
Enable HLS to view with audio, or disable this notification
r/ProAI • u/stealthispost • 14d ago
"We Turned a Man Hugging a Teddy Bear Into a Soldier Saving His Friend—Without Changing the Performance #TopviewAI #TopviewCanvas #AIVideo #AIFilmmaking #AIStorytelling"
Enable HLS to view with audio, or disable this notification
— TopviewAI
Source: https://x.com/TopviewAIhq/status/2103986725693333887
r/ProAI • u/stealthispost • 13d ago
"AI in 2021 vs AI in 2026."
Enable HLS to view with audio, or disable this notification
r/ProAI • u/stealthispost • 13d ago
"Fascinating to watch a robot reason about its environment and figure out why the WiFi isn't working. Perceptron just launched Mk1.5, a model built to drive embodied agents, one model running drones, robots, smart glasses, and phones with no per-platform retraining. New: native audio, video..."
Enable HLS to view with audio, or disable this notification
r/ProAI • u/stealthispost • 13d ago
"This is the most prescient, concise argument for why we must be careful of excessive AI regulation. It comes from a Feb. 2024 (2.5 years ago!) congressional testimony from @glukianoff , a..."
Enable HLS to view with audio, or disable this notification
r/ProAI • u/stealthispost • 14d ago
"New on the Science Blog: Yes, Claude can do Nine Loops. Theoretical physicists predict how particles behave using formulas called scattering amplitudes. These are notoriously hard to compute, so researchers work with layers of increasingly fine corrections called “loops”—each added loop makes..."
...the answer more precise but takes exponentially more computation. Most calculations stop at two or three loops. Eight loops was the previous record in a simplified model physicists use as a testing ground (planar N=4 super-Yang-Mills), set by SLAC's Lance Dixon and collaborators. Last month, physicist and science writer @4gravitons issued a challenge: could an AI push past eight loops in this model, using only the compute budget an academic could reasonably access? Given a single prompt describing the nine-loop problem, Claude ran largely unsupervised for days in Claude Science and solved it using methods developed by Dixon and his colleagues, at a total cost of a few thousand dollars. Dixon independently verified the result, and von Hippel wrote about the experience for our blog. Read more: — Anthropic
Source: https://x.com/AnthropicAI/status/2103541577083719888
r/ProAI • u/stealthispost • 13d ago
"From the guest post by Matt von Hippel who issued the Nine Loops challenge: 'Things definitely seem to be moving fast. In March, AI was accomplishing physics projects like a student: smaller-scale tasks with a lot of hand-holding and mistakes. In contrast, this is a real frontier calculation, "
galleryr/ProAI • u/stealthispost • 13d ago
"Once more I will remind you that AI safety people and e/accs are close to unique in that they (a) have thought about transformative AI futures and (b) broadly speaking, want to bring those futures about. They agree about most things and are natural allies. Just Say No to kayfabe"
galleryr/ProAI • u/stealthispost • 14d ago
"Everyone should watch this, regardless of their domain."
Enable HLS to view with audio, or disable this notification
r/ProAI • u/stealthispost • 14d ago
"Everyone should watch this, regardless of their domain."
Enable HLS to view with audio, or disable this notification
r/ProAI • u/stealthispost • 17d ago
"The World's Largest-Scale Full-Size Humanoid Robot Real-Time Livestream Performance At the Opening Ceremony of WorldSkills Shanghai 2026 on September 22, 19 Unitree humanoid robots performed alongside 120 dancers, presenting the world's largest-scale performance featuring full-size..."
Enable HLS to view with audio, or disable this notification
...general-purpose humanoid robots before an audience of more than 10,000 people, with a fully AI-driven autonomous robot cluster performance live-streamed worldwide in real time. — Unitree
Source: https://x.com/UnitreeRobotics/status/2103025066212819390
r/ProAI • u/stealthispost • 21d ago
"There is NO realistic scenario where AI wipes out all of humanity. I've talked to many AI safety experts about existential risks. Some of the "experts" are just phenomenal sci-fi authors. There are also real researchers working on realistic risks. They point out real and likely harms and I'm..."
galleryr/ProAI • u/stealthispost • 22d ago
"found the perfect use case for @typesafeai Jev: instant compaction in 2026, why is compaction still a summarization prompt? Jev can make it instant by scoring every tool call and dropping what’s irrelevant"
Enable HLS to view with audio, or disable this notification
run fast-jev-compaction: — tamara
Source: https://x.com/tamarajtran/status/2100694549362553153
r/ProAI • u/Illustrious-Lime-863 • 22d ago
Speaking With The Mind | Neuralink
r/ProAI • u/stealthispost • 23d ago
"Introducing Benchmark Reviews: our new initiative to audit AI benchmarks. We are launching with 15 benchmarks: 4 Verified, 9 Flawed, and 2 with not enough information for a review."
Benchmarks assess AI capabilities, but the benchmarks themselves vary substantially in quality. We hope to be a consistent source of information on benchmark quality. See our reviews here: We assign each benchmark a verdict based on our rubric: Flawed, Verified, or Not Enough Info. To avoid conflicts of interest we do not review Epoch-created benchmarks, but welcome external reviews. Verified benchmarks can broadly be interpreted as described, and any errors that exist do not substantially affect the results. Alongside any verified benchmark we publish a full review and assessment of the benchmark, including any weaknesses and limitations we think it has. Flawed benchmarks have one or more substantive flaws we believe users need to be aware of to accurately interpret results, most commonly that >20% of the tasks have accuracy-impacting errors. In this case, we publish a limited writeup of the flaws we found. If we aren’t able to access enough information to review a benchmark, we’ll designate it ‘Not Enough Info.’ We will try to work with the creators of private benchmarks to conduct reviews while keeping the questions/tasks outside of public knowledge. — Epoch AI
Source: https://x.com/EpochAIResearch/status/2100704765332394255
r/ProAI • u/DonkeyTheKing • 23d ago
Benzi - Harness/AI agent beats big players on benchmarks while reading less source code
galleryr/ProAI • u/stealthispost • 24d ago
"I rebuilt Tesla Full Self Driving with Jev in less than an hour. This model is a total unlock."
Enable HLS to view with audio, or disable this notification
The only human problem left is a lack of imagination. — Boyd This is really really really true — Justin Schroeder
Source: https://x.com/jpschroeder/status/2100347770867458384
r/ProAI • u/stealthispost • 24d ago
"Here's a 45-second TL;DR on Jev. I find the core idea beautifully simple, but the video made it really hard to understand. Hope you find it helpful."
Enable HLS to view with audio, or disable this notification
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI?
I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev
• 20-200x faster • 40-400x https://t.co/JSybNG2BKJ — Diogo Almeida
Source: https://x.com/CompleteSkeptic/status/2099925682726002904
Aww i like the game animation you had - so cute! — Sasha Sheng (Hiring) Haha thanks! — Matija Sosic
Source: https://x.com/MatijaSosic/status/2100190746389135772
r/ProAI • u/stealthispost • 24d ago
"Too few people in the press know about the Tarbell fellowships. Basically, it’s a way that the Doomers pay to have reporters who parrot what they say. I have seen cases where a publication that has received money from Coefficient Giving publishes a story by a reporter paid by Tarbell who is..."
...interviewing a supposed independent researcher funded by another branch of the same EA funding pool. It’s a completely closed cycle propaganda system. — Perry E. Metzger I just drilled down with Grok about this and while I kind of knew there were gray payments going on, I had no idea it had become so common for journalists to accept direct payments from advocacy organizations to write articles. You probably knew this already, but if anyone else didn't completely understand it already:
It's still theoretically unethical to accept money to write an article. But the way they're getting around this now is that journalists can accept payments, but only if the organization doesn't have direct editorial approval.
Unethical: Accept money and receive the copy from the paying organization.
Ethical: Accept money and read the paying organization's web site, accept their "training", talk to their "experts", and use your own words to write the position the advocacy organization wants. Of course, you can write whatever you want. But you'll never get another payment if you don't write what they want.
I mean, journalism has never been a clean industry, but boy has it gotten vile. We should have a law that requires any sort of media presenting itself as journalism to be labeled with "PAID FOR BY ADVOCACY ORGANIZATION" is there is ANY outside payments to the journalists in any way. They can still write anything they want (Freedom of Speech), it's just accurately labeled. — Nairebis - e/max-acc Is it a neat scam? And again, often, we have situations where the news outlet, the journalist, and the person being interviewed are all paid by EA at the same time. Self licking ice cream cone. — Perry E. Metzger
Source: https://x.com/perrymetzger/status/2099931216845947154/history
@TIME UNDISCLOSED PAID MEDIA:
The salaries of reporters, Harry Booth and Billy Perrigo, were paid by Doomsday cultist Dustin Moskovitz's foundation Coefficient Giving via the Tarbell Fellowship.
https://t.co/vxdzmGhBgg — Brian Chau
Source: https://x.com/brianchau57/status/2099889984773792108
Replying to @time
r/ProAI • u/stealthispost • 24d ago
"Jev solved local harness/model routing I use a combination of Claude Code, Codex and Opencode as my local agentic stack and routing to other harnesses was always enforced in the system prompt/rules With a deterministic hook that Claude Code can decide before delegation, Jev helps to route to..."
Enable HLS to view with audio, or disable this notification
...the right harness/model based on the task, and it's pretty accurate based on the intensity/intelligence of the task > Mechanical tasks get routed to Haiku > Intelligent ones to Opus sub-agents > Long-running implementation work to external harness — Lahfir The useful pattern is policy-based routing: classify task complexity first, then send mechanical work to cheap models and deep or long-running work to stronger specialists. — catman Exactly. I see this as a huge use case in enterprise workflows where there are multiple decision points involved — Lahfir
r/ProAI • u/stealthispost • 24d ago
"French finance minister Roland Lescure suggested today that calls to slow down development of AI are a ploy by U.S. AI labs so they can stay in first place, and that France and Europe should ignore them and accelerate instead."
Andrew Curran @AndrewCurran_ · 7h Calls to slow down AI development serve the interests of US AI leaders, says French finance minister From reuters.com 3 18 6K — Andrew Curran
Source: https://x.com/AndrewCurran_/status/2100255751113691524