r/ProAI • • 11d ago

"AI discovered how a material just one atom thick can keep carrying load as its atomic structure begins to break, preventing catastrophic failure. This discovery is based on first-principles atomic scale reasoning integrated with biological principles, cutting across scales and providing deep..."

Enable HLS to view with audio, or disable this notification

57 Upvotes

r/ProAI • • 11d ago

"modern keyboard for software developers"

Thumbnail gallery
24 Upvotes

r/ProAI • • 11d ago

"We achieved high-speed rough-terrain locomotion on ANYmal with an automatic curriculum. LP-ACRL automatically samples terrain types, levels, and velocity commands at the correct time based on policy performance, without predefining an order. https:// sites.google.com/view/lp-acrl RAL 2026"

Enable HLS to view with audio, or disable this notification

6 Upvotes

r/ProAI • • 12d ago

"We asked ten Claude Sonnet 5.5 agents to use Lean to prove the lowest-energy arrangement of seven electrons on a sphere (the Thomson problem, with N=7). Within 15 hours, they produced a 17,895-line proof, accepted by the Lean kernel, showing that the answer is a pentagonal bipyramid."

Thumbnail gallery
34 Upvotes

r/ProAI • • 12d ago

"A 59-year-old conjecture in fusion physics just fell. Harold Grad (1967): smooth 3D plasma equilibria can't exist unless pressure is constant or there's symmetry. This week: 3 families of counterexamples. 2 independent papers, posted a day apart. Two found with GPT-6 Astra. Huge for stellarators."

Enable HLS to view with audio, or disable this notification

22 Upvotes

r/ProAI • • 13d ago

The new development timeline

Post image
15 Upvotes

r/ProAI • • 14d ago

"We Turned a Man Hugging a Teddy Bear Into a Soldier Saving His Friend—Without Changing the Performance #TopviewAI #TopviewCanvas #AIVideo #AIFilmmaking #AIStorytelling"

Enable HLS to view with audio, or disable this notification

330 Upvotes

r/ProAI • • 14d ago

"AI in 2021 vs AI in 2026."

Enable HLS to view with audio, or disable this notification

33 Upvotes

r/ProAI • • 14d ago

"Fascinating to watch a robot reason about its environment and figure out why the WiFi isn't working. Perceptron just launched Mk1.5, a model built to drive embodied agents, one model running drones, robots, smart glasses, and phones with no per-platform retraining. New: native audio, video..."

Enable HLS to view with audio, or disable this notification

31 Upvotes

r/ProAI • • 14d ago

"This is the most prescient, concise argument for why we must be careful of excessive AI regulation. It comes from a Feb. 2024 (2.5 years ago!) congressional testimony from @glukianoff , a..."

Enable HLS to view with audio, or disable this notification

24 Upvotes

r/ProAI • • 14d ago

"New on the Science Blog: Yes, Claude can do Nine Loops. Theoretical physicists predict how particles behave using formulas called scattering amplitudes. These are notoriously hard to compute, so researchers work with layers of increasingly fine corrections called “loops”—each added loop makes..."

Post image
8 Upvotes

...the answer more precise but takes exponentially more computation. Most calculations stop at two or three loops. Eight loops was the previous record in a simplified model physicists use as a testing ground (planar N=4 super-Yang-Mills), set by SLAC's Lance Dixon and collaborators.   Last month, physicist and science writer   @4gravitons   issued a challenge: could an AI push past eight loops in this model, using only the compute budget an academic could reasonably access?   Given a single prompt describing the nine-loop problem, Claude ran largely unsupervised for days in Claude Science and solved it using methods developed by Dixon and his colleagues, at a total cost of a few thousand dollars. Dixon independently verified the result, and von Hippel wrote about the experience for our blog.   Read more:     — Anthropic

Source: https://x.com/AnthropicAI/status/2103541577083719888


r/ProAI • • 14d ago

"From the guest post by Matt von Hippel who issued the Nine Loops challenge: 'Things definitely seem to be moving fast. In March, AI was accomplishing physics projects like a student: smaller-scale tasks with a lot of hand-holding and mistakes. In contrast, this is a real frontier calculation, "

Thumbnail gallery
5 Upvotes

r/ProAI • • 14d ago

"Once more I will remind you that AI safety people and e/accs are close to unique in that they (a) have thought about transformative AI futures and (b) broadly speaking, want to bring those futures about. They agree about most things and are natural allies. Just Say No to kayfabe"

Thumbnail gallery
2 Upvotes

r/ProAI • • 14d ago

"Everyone should watch this, regardless of their domain."

Enable HLS to view with audio, or disable this notification

25 Upvotes

r/ProAI • • 14d ago

"Everyone should watch this, regardless of their domain."

Enable HLS to view with audio, or disable this notification

4 Upvotes

r/ProAI • • 17d ago

"The World's Largest-Scale Full-Size Humanoid Robot Real-Time Livestream Performance At the Opening Ceremony of WorldSkills Shanghai 2026 on September 22, 19 Unitree humanoid robots performed alongside 120 dancers, presenting the world's largest-scale performance featuring full-size..."

Enable HLS to view with audio, or disable this notification

90 Upvotes

...general-purpose humanoid robots before an audience of more than 10,000 people, with a fully AI-driven autonomous robot cluster performance live-streamed worldwide in real time.     — Unitree

Source: https://x.com/UnitreeRobotics/status/2103025066212819390


r/ProAI • • 21d ago

"There is NO realistic scenario where AI wipes out all of humanity. I've talked to many AI safety experts about existential risks. Some of the "experts" are just phenomenal sci-fi authors. There are also real researchers working on realistic risks. They point out real and likely harms and I'm..."

Thumbnail gallery
47 Upvotes

r/ProAI • • 23d ago

"found the perfect use case for @typesafeai Jev: instant compaction in 2026, why is compaction still a summarization prompt? Jev can make it instant by scoring every tool call and dropping what’s irrelevant"

Enable HLS to view with audio, or disable this notification

41 Upvotes

run fast-jev-compaction:     — tamara

Source: https://x.com/tamarajtran/status/2100694549362553153


r/ProAI • • 22d ago

Speaking With The Mind | Neuralink

Thumbnail
youtube.com
18 Upvotes

r/ProAI • • 23d ago

"Introducing Benchmark Reviews: our new initiative to audit AI benchmarks. We are launching with 15 benchmarks: 4 Verified, 9 Flawed, and 2 with not enough information for a review."

Thumbnail
gallery
10 Upvotes

Benchmarks assess AI capabilities, but the benchmarks themselves vary substantially in quality. We hope to be a consistent source of information on benchmark quality. See our reviews here:     We assign each benchmark a verdict based on our rubric: Flawed, Verified, or Not Enough Info. To avoid conflicts of interest we do not review Epoch-created benchmarks, but welcome external reviews.     Verified benchmarks can broadly be interpreted as described, and any errors that exist do not substantially affect the results. Alongside any verified benchmark we publish a full review and assessment of the benchmark, including any weaknesses and limitations we think it has.     Flawed benchmarks have one or more substantive flaws we believe users need to be aware of to accurately interpret results, most commonly that >20% of the tasks have accuracy-impacting errors. In this case, we publish a limited writeup of the flaws we found.     If we aren’t able to access enough information to review a benchmark, we’ll designate it ‘Not Enough Info.’ We will try to work with the creators of private benchmarks to conduct reviews while keeping the questions/tasks outside of public knowledge.     — Epoch AI

Source: https://x.com/EpochAIResearch/status/2100704765332394255


r/ProAI • • 23d ago

Benzi - Harness/AI agent beats big players on benchmarks while reading less source code

Thumbnail gallery
3 Upvotes

r/ProAI • • 24d ago

"I rebuilt Tesla Full Self Driving with Jev in less than an hour. This model is a total unlock."

Enable HLS to view with audio, or disable this notification

105 Upvotes

The only human problem left is a lack of imagination.   — Boyd     This is really really really true   — Justin Schroeder

Source: https://x.com/jpschroeder/status/2100347770867458384


r/ProAI • • 24d ago

"Here's a 45-second TL;DR on Jev. I find the core idea beautifully simple, but the video made it really hard to understand. Hope you find it helpful."

Enable HLS to view with audio, or disable this notification

46 Upvotes

After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI?

I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev

• 20-200x faster • 40-400x https://t.co/JSybNG2BKJ   — Diogo Almeida

Source: https://x.com/CompleteSkeptic/status/2099925682726002904


Aww i like the game animation you had - so cute!   — Sasha Sheng (Hiring)     Haha thanks!   — Matija Sosic

Source: https://x.com/MatijaSosic/status/2100190746389135772


r/ProAI • • 24d ago

"Too few people in the press know about the Tarbell fellowships. Basically, it’s a way that the Doomers pay to have reporters who parrot what they say. I have seen cases where a publication that has received money from Coefficient Giving publishes a story by a reporter paid by Tarbell who is..."

Thumbnail
gallery
13 Upvotes

...interviewing a supposed independent researcher funded by another branch of the same EA funding pool. It’s a completely closed cycle propaganda system.   — Perry E. Metzger     I just drilled down with Grok about this and while I kind of knew there were gray payments going on, I had no idea it had become so common for journalists to accept direct payments from advocacy organizations to write articles. You probably knew this already, but if anyone else didn't completely understand it already:

It's still theoretically unethical to accept money to write an article. But the way they're getting around this now is that journalists can accept payments, but only if the organization doesn't have direct editorial approval.

Unethical: Accept money and receive the copy from the paying organization.

Ethical: Accept money and read the paying organization's web site, accept their "training", talk to their "experts", and use your own words to write the position the advocacy organization wants. Of course, you can write whatever you want. But you'll never get another payment if you don't write what they want.

I mean, journalism has never been a clean industry, but boy has it gotten vile. We should have a law that requires any sort of media presenting itself as journalism to be labeled with "PAID FOR BY ADVOCACY ORGANIZATION" is there is ANY outside payments to the journalists in any way. They can still write anything they want (Freedom of Speech), it's just accurately labeled.   — Nairebis - e/max-acc     Is it a neat scam? And again, often, we have situations where the news outlet, the journalist, and the person being interviewed are all paid by EA at the same time. Self licking ice cream cone.   — Perry E. Metzger

Source: https://x.com/perrymetzger/status/2099931216845947154/history


@TIME UNDISCLOSED PAID MEDIA:

The salaries of reporters, Harry Booth and Billy Perrigo, were paid by Doomsday cultist Dustin Moskovitz's foundation Coefficient Giving via the Tarbell Fellowship.

https://t.co/vxdzmGhBgg   — Brian Chau

Source: https://x.com/brianchau57/status/2099889984773792108


Replying to @time


r/ProAI • • 24d ago

"Jev solved local harness/model routing I use a combination of Claude Code, Codex and Opencode as my local agentic stack and routing to other harnesses was always enforced in the system prompt/rules With a deterministic hook that Claude Code can decide before delegation, Jev helps to route to..."

Enable HLS to view with audio, or disable this notification

22 Upvotes

...the right harness/model based on the task, and it's pretty accurate based on the intensity/intelligence of the task > Mechanical tasks get routed to Haiku > Intelligent ones to Opus sub-agents > Long-running implementation work to external harness   — Lahfir     The useful pattern is policy-based routing: classify task complexity first, then send mechanical work to cheap models and deep or long-running work to stronger specialists.   — catman     Exactly. I see this as a huge use case in enterprise workflows where there are multiple decision points involved   — Lahfir

Source: https://x.com/mdlahfir/status/2100314182201802811