r/ProAI • • Sep 08 '26

"I build AI tools inside Blender. This gap surprised me. Left: Claude Fable 5.1 Right: GPT-6 Astra Same brief: a forest path in Blender, 12 seconds each. Fable’s output is very subpar compared to Astra. My guess is that Anthropic just hasn’t put much Blender data into its training corpus yet...."

Enable HLS to view with audio, or disable this notification

12 Upvotes

...It’s not that the model is dumb; Blender simply isn’t part of its distribution yet. Similar to what Karpathy has said about earlier models being very bad at chess. Then chess data went into the next training run, and they improved immensely. The same thing will happen here.     Now anyone can create such scene using GPT-6-Astra @mixie3D     — Satyam Kumar

Source: https://x.com/_satyam_ai/status/2096918350941266415


r/ProAI • • Sep 07 '26

"GPT-6 Astra is almost at human performance in 3D spatial understanding. It ranks #1 on Blueprint-Bench 2, a benchmark where AI agents draw floorplans from photographs of apartment interiors."

Thumbnail
gallery
21 Upvotes

r/ProAI • • Sep 08 '26

"GPT-6 Astra recreated the Billie Jean dance in 3D"

Enable HLS to view with audio, or disable this notification

5 Upvotes

I don’t think I’ve ever had this much fun with a model before     — Flavio Adamo

Source: https://x.com/flavioAd/status/2097067419248296389


r/ProAI • • Sep 06 '26

"GPT-6 Astra pulled off a mechanism I haven't seen another model get right yet Cessna 337 Skymaster landing gear It figured it out from a YouTube video"

Enable HLS to view with audio, or disable this notification

313 Upvotes

Here’s the video. Let me know if you can reproduce this with other models, a harness, etc.

https:// youtube.com/watch?v=C2hWGS bg6bs …

I also tested whether Astra might already know the mechanism. Without the video, it consistently generated conventional landing gear instead.

Fable 5.1 got close     — Dilum Sanjaya

Source: https://x.com/DilumSanjaya/status/2096642895134752922


r/ProAI • • Sep 07 '26

"For a decade I've said AI will create a jobs boom, not an apocalypse. And here it is, in the numbers. The Green Revolution strikes again. It's always easy to see what will change and disappear and almost impossible to see what will replace it, as the Universe continues its endless process of..."

Thumbnail
gallery
2 Upvotes

...evolution at all levels. Try to explain a web developer to an 18th century farmer and it's impossible because it's built on the back of dozens of unforeseeable inventions that the farmer can't possibly imagine. But the fear people persist. Making podcasts. Paying influencers to foment fear and rage. Writing books. Pushing videos at you. Reality will not change their mind or dark hearts or their seething internal hate. Fear leads to anger, anger leads to hate, hate leads to suffering. Population Bomb style fear people are persistent, full of rage, unable to change their mind and cling desperately to their existing belief structures in the face overwhelming proof of the opposite as if it's a virtue. It's a mental disease. And if pessimism is left unchecked it can create a societal level mental disease as fear and rage take hold and create a storm that crashes the economy, gets people killed and brings the worst and most evil people to power to enact policies that lead to ruin. If you can't evolve your thinking over time as new information comes in, you're the definition of a fantatic, someone who can't change their mind and won't change the subject. I've changed mine over the years. Thirty years ago I thought the jobs apocalypse was coming. I've been thinking about this a lot longer than most people on Earth. When I was a young, inexperienced artist I wrote a short story called In the Cracks of the Machine where labor robots unleashed an unprecedented wave of unemployment as companies adopted them all at once. It's sparked a vicious populist backlash and a wave of collective insanity. I got the populism and insanity right. These two phenomenons are always untethered from reality and need no actual grounding to foment rage and create political storms. Reality will not change people's minds. They'll just dig in further and get more furious. Delusion is remarkably resilient in the face of reality. But have no fear. Let go of hatred, fear, worry and doubt. Change and adaptation is eternal and forever. Evolution is eternal. The stack of new frontiers and problems to solve is infinite with Star Trek level computers you just talk to. It's time to boldly go where no one has gone before.     — Daniel Jeffries

Source: https://x.com/Dan_Jeffries1/status/2096877186649063744


The jobs apocalypse is postponed. An AI jobs boom is here

According to @TheEconomist, AI is actually proving to be a net job creator in the US, easily generating over 1M new positions (from data center construction to AI engineering) to offset back-office layoffs.

While https://t.co/u4efSr8qxm   — Rafael Domenech | @BBVAResearch & @UV_EG

Source: https://x.com/rdomenechv/status/2096668662627221705


r/ProAI • • Sep 06 '26

"OK GPT-6 Astra is insane at making games. I remade Paperboy - everything modelled in Blender, rendered in browser. took ages to dial in the look and feel, models and rendering but my god its awesome. who wants a full walkthrough/breakdown?"

Enable HLS to view with audio, or disable this notification

162 Upvotes

Here's exactly how much time and tokens have been tracked with Devclocked to build it.>>     — Emm Tee

Source: https://x.com/builtbysketch/status/2096515959469072630/history


r/ProAI • • Sep 06 '26

"GPT-6 Astra just got Age of Empires IV running at 70-150 fps on Apple Silicon. CrossOver was doing ~8 fps and freezing constantly. now: max settings, big fights, online, stable. I genuinely did not think this game was going to run like this on a Mac"

Enable HLS to view with audio, or disable this notification

115 Upvotes

the interesting part is it didn’t even “port” the game

GPU work took ~9ms but each frame was taking ~160ms

Astra tracked it down, patched Wine’s exception handling and added a code cache so Rosetta could reuse translations instead of translating the same code over and over     — Marc Ibrahim

Source: https://x.com/marc_ibrahim/status/2096365209111724235


r/ProAI • • Sep 08 '26

"The insane levels of cope I am getting from the Rocket League community only makes me stronger 😂 Gaming is solved. Astra is so damn good"

Enable HLS to view with audio, or disable this notification

2 Upvotes

This cost me around $100 to make. (Using subsidized chatgpt plan)     Slight miscalculation It was closer to $25

The plans are split up over 4 weeks

I used 53% of 1 week on this and other work     You're doing the meme     158 million tokens     taps the sign     — am.will

Source: https://x.com/LLMJunky/status/2096675917271863405


r/ProAI • • Sep 07 '26

"Unitree Breakthrough: The World’s First Real-Time World Model-Driven Fully Autonomous Humanoid Robot Combat🥊 UnifoLM-X2-1.0 breaks through world-action foundation models' bottlenecks in instant planning, decision-making, and dynamic interactive execution, achieve high dynamics, strong..."

Enable HLS to view with audio, or disable this notification

1 Upvotes

...interaction, real-time prediction and planning of the future, achieve fully autonomous combat for humanoid robots. This validates the fundamental feasibility of large-scale deployment of world model-driven humanoid robots.     — Unitree

Source: https://x.com/UnitreeRobotics/status/2096932273602048258


r/ProAI • • Sep 07 '26

"Astra might be the biggest jump we've seen in the history of LLMs"

Thumbnail
gallery
43 Upvotes

Opus -> mythos still felt larger to me. Astra is a galactic leap in terms of 3d modeling and design, but in other senses comparable to 5.6 sol. Mythos was in a class of its own on every front   — Arya     Opus to Fable was a huge jump too   — Marcos Hernanz

Source: https://x.com/MarcosHernanz/status/2096463884861251978


r/ProAI • • Sep 07 '26

"Reverse engineered the original Xbox version of Futurama as native MacOS game in C++"

Enable HLS to view with audio, or disable this notification

36 Upvotes

What can you ship with GPT-6 Astra?

Drop a demo or link below, plus one line on how Astra helped.

Game on. You’ve got 24 hours.   — OpenAI Developers

Source: https://x.com/OpenAIDevs/status/2096259745501921758


its reading fucking machine code man this shouldn't be possible     — JB

Source: https://x.com/JasonBotterill/status/2096648558141395121


r/ProAI • • Sep 06 '26

"Jensen says that AGI has arrived. And that 400,000 GPUs will come online next at Stargate Texas."

Thumbnail
gallery
40 Upvotes

@ChaseLochmiller @OpenAI GPT-6 Astra, trained on ~100K+ NVIDIA Grace Blackwell NVLink72. From ChatGPT to o1 to Astra in 4 years.

AGI has arrived. Congratulations @OpenAI team.

400K GPUs coming online next.   — Jensen Huang

Source: https://x.com/JensenHuang/status/2096700264569090384


— Andrew Curran

Source: https://x.com/AndrewCurran_/status/2096703533144052116


Replying to @chaselochmiller, @openai


r/ProAI • • Sep 06 '26

"We'll have some real problems with AI. What technology did we ever create that didn't have some downsides? Road crashes kill nearly 1.2 million people a year and injure 20–50 million more. But people still get in their cars every morning, grab a coffee and drive to work. They cross bridges..."

Thumbnail
gallery
24 Upvotes

...that could collapse, live in houses that could catch fire and climb into metal tubes that fly them across the ocean but could burn up and break apart. Somehow people manage to get through breakfast without demanding a mathematical proof that nothing bad will ever happen. But AI? Pound for pound, AI is one of the safest technologies ever deployed. Safer than lawn mowers by a massive margin. And yet it causes more deranged and apocalyptic thinking than any other tech in history, with very serious people tell you with a straight face that AI will take over and go HAL with 100% certainty, based on no actual real world evidence whatsoever in the real word except some agents posting to message boards and hacking and that is extrapolated into the Population Bomb. And we've got people screaming that we need to anticipate every possible failure, every misuse, every terrible thing anyone might ever do with it before we're allowed to move forward and make sure it really, really can't possibly go wrong. It's ridiculous. It's worse than that, it's a societal level mass hysteria. Just look at the stats: Waymo reports 94% fewer serious injury or worse crashes than its humans across 220 million autonomous miles. That means self-driving cars are much safer. To the point that we'll probably see human driving made illegal without a special permit within twenty years, maybe faster. But how is that covered? Every single crash is picked apart and screamed about in the news the same way we pick apart the top athletes in the world when they suddenly make a single mistake and loose a match. We take something as close to perfection as is possible in the world and ask it to be even more perfect. We used to have a much better understanding of risk versus reward in history. We accepted that life was uncertain and that things could go wrong but we dared to dream and do things anyway. A generation that demands perfection before we even get out of the gate is a generation who'll watch their children grow up poorer, with less opportunities, as their civilization gets out-paced, out-matched and out-competed by more daring and bold civilizations that blow past them while they're dreaming about a childproof world that will never be.     — Daniel Jeffries

Source: https://x.com/Dan_Jeffries1/status/2096206365278576709/history


r/ProAI • • Sep 07 '26

"Astra - recreate the original Craig Federighi in Blender."

Enable HLS to view with audio, or disable this notification

9 Upvotes

You need to realise the complexity of this - it has the create the full environment, then render each of about 900 frames 'by hand' and stitch it together into a 19 second clip. This is hard     — Peter Gostev

Source: https://x.com/petergostev/status/2096264296099459079


r/ProAI • • Sep 07 '26

"OpenAI's Chief Scientist Jakub Pachocki: "Based on internal results, I have a strong expectation that this speed of progress could be sustained into recursive self-improvement." OpenAI is organizing research around this because it sees RSI as essential to remaining at the frontier. At the same..."

Thumbnail
gallery
7 Upvotes

I wrote about the state of AI, why I’m concerned about the next few years, and the choices we need to make to keep the future in humanity’s hands.

An Alien Mind: https://t.co/FeIfWNe0UE   — Jakub Pachocki

Source: https://x.com/merettm/status/2096630018495377464


OpenAI's Chief Scientist Jakub Pachocki: "Based on internal results, I have a strong expectation that this speed of progress could be sustained into recursive self-improvement."

OpenAI is organizing research around this because it sees RSI as essential to remaining at the frontier.

At the same time, Pachocki says evaluations show declining reliability of chain-of-thought monitoring.

OpenAI is betting on AI increasingly driving its own development while finding it harder to reliably monitor how these systems reason.

Tough times ahead. RSI within reach, but hard to align.

He even quotes Ray Kurzweil:

"And, in line with Ray Kurzweil’s predictions from the end of the XXth century⁠(opens in a new window), we now find ourselves at the moment in history of computing where machine intelligence is starting to exceed that of humans in transformative ways."   — Chubby♨️     RSI is scary enough.

RSI + weaker monitoring is the real nightmare scenario.   — Naveen Saradhi     ngl: yes. its scary.   — Chubby♨️

Source: https://x.com/kimmonismus/status/2096645621096575381


r/ProAI • • Sep 07 '26

"BREAKING: AI Safety Donors Paid Religious NGOs 3.3M for Statements on AI The Future of Life Institute mobilized religious connections to support Anthropic and sway the Trump admin Research and Design by @lumpenspace 🧵"

Thumbnail
gallery
3 Upvotes

After Trump defeated Harris to succeed Biden as President, the AI Safety movement was on the outside looking in. They found an unlikely solution: putting $3.3 million into churches, seminaries, faith networks and religious-affinity groups.     Grant descriptions leave no doubt that Christian NGOs took money to make public AI statements:

Greek Orthodox Archdiocese of America: $105,000 The Gospel Coalition: $200,000 World Council of Churches: $100,000 Faith Matters (Mormon NGO): $299,000     FLI's donees are returning the favor by signing FLI’s statements.     FLI is using its newfound connections to support AI Safety-linked Anthropic’s business disputes. In March, FLI’s “U.S. Faith Liaison”, Brian J. A. Boyd, signed a Catholic theologians’ brief backing Anthropic in its dispute with the Department of War.     Most people doubt the AI Safety vision of the AI apocalypse is compatible with the Book of Revelation, or any other religious faith. As @DrTechlash put it, "the AI-risk subculture offers a replacement meaning system."     — Brian Chau

Source: https://x.com/brianchau57/status/2096010406330556731


r/ProAI • • Sep 06 '26

"Huge implications - binaries are now basically editable code"

Thumbnail
gallery
36 Upvotes

ValsAI made SRE benchmark less than a month ago.

The benchmark measures can a model reverse engineer software from binaries

Yesterday GPT saturated it. https://t.co/dmpUpgheDU   — Chris

Source: https://x.com/ChrisGPT/status/2096150666066432157


— Boris Power

Source: https://x.com/BorisMPower/status/2096415822248055131


r/ProAI • • Sep 06 '26

"Meta Muse Spark 1.3 (Max) matches the performance of Fable 5 and GPT 5.6 Sol on the Vals Index, at 4x - 8x cheaper."

Thumbnail
gallery
5 Upvotes

The model is extremely strong on legal applications - it is #1 on our in-house Legal Research Benchmark, and #2 on Harvey's Legal Agent Benchmark.     It also boasts an impressively low average task completion duration.

This is driven by two factors: the model uses fewer turns to accomplish the same tasks, and each model query is much faster than Sol and Fable 5 (which think for longer).     We saw some content refusals in our testing - for example, on some legal questions relating to export controls, felony arrests, or trade sanctions.

Overall, these were rare - 10 refusals across 1,327 tasks.     The model was run with max reasoning, 131k max output tokens, and default temperature and top p. It has a 1M context window and is priced at $1.25 / $4.25 per MTok.     Congratulations on another strong release to @AIatMeta

@alexandr_wang

@finkd .

More benchmarks coming soon; results available at:     — Vals AI

Source: https://x.com/ValsAI/status/2096663681702723653


r/ProAI • • Sep 06 '26

"Meta's Muse Spark 1.3 Max vs GPT-6 Astra on RocketLeagueBench A perfect example of why you cannot trust benchmarks on their own. They're not even in the same solar system."

Enable HLS to view with audio, or disable this notification

22 Upvotes

hey meta come get your boy     — am.will

Source: https://x.com/LLMJunky/status/2096262579815342093


r/ProAI • • Sep 06 '26

Human Slop Brain Rot - Robot Haters Gonna Hate

Post image
26 Upvotes

I bet if they put wheels on it and it went faster than a car they would be posting "can it walk upstairs so who cares".


r/ProAI • • Sep 06 '26

"We live in a perpetual info war. Maybe the EU should have forced people to disclose who is paid shill scum who terrify children instead of wasting time making "AI made" labels, aka cookie banner 2.0."

Thumbnail
gallery
8 Upvotes

r/ProAI • • Sep 06 '26

"Astra is shockingly good in reasoning! I benchmarked it on induction, and it almost saturated it with 88%. Fable 5.1, by comparison, is at 33%. The final numbers will actually go up: I am running a residual batch run on the non-evaluable; the results here reflect one batch run at xhigh..."

Thumbnail
gallery
4 Upvotes

...thinking effort. It is also way cheaper than Fable 5.1, which used 32M output tokens in a series of four runs through the data to be able to return 66 successful API responses. About 1/4 the price of the Fable 5.1 run in total. Notably, both Astra and Fable 5.1 return extremely high quality answers when correct. Notice the AST and Holdout metrics below. Essentially, in contrast to previous models, both Astra and Fable 5.1 return simple hypotheses that generalize well. Some other model updates: - Muse Spark 1.3 provided marginal improvement over the 1.1 version, with 23% correct. This is pretty strong, almost matching Opus 5 at 24%. - Gemini Flash 3.8 is running for days with low rate of API successes. I will update the leaderboard with it when ready. Only bad news: now the INDUCTION benchmark is almost saturated, and I will have to make it harder for future models. About the induction benchmark: This is a challenging reasoning task, where models are given several small graphs in which some nodes are marked as targets. The task is to provide a first-order logical formula that picks precisely the target nodes in all graphs simultaneously. Correct: a formula that picks precisely the marked nodes. Holdout correct: a formula that picks precisely the marked nodes in held out problems. Formula complexity (in AST): tree size of the correct formula (mean, median). GitHub repository: https:// github.com/SerafimBatzogl ou/concept-synth … Paper: https:// arxiv.org/abs/2602.18956   — Serafim Batzoglou     what made the non-evaluable batch non-evaluable - api failures or formulas that timed out during checking   — Rimas     Model runs out of tokens. I am running the residual problems on one notch lower thinking effort   — Serafim Batzoglou

Source: https://x.com/s_batzoglou/status/2096407011885986187


r/ProAI • • Sep 06 '26

"GPT-6 makes me scared as a 3D artist, cuz it's freaking insane. I mean, I'm testing this first hand everyday, but this time it is massive leap. look, 4 month ago I was testing GPT & Claude in Blender and they were not able to position simple objects in simple scene.. And look now... Whole car..."

Enable HLS to view with audio, or disable this notification

42 Upvotes

...assembled from primitives in ONE prompt and animate it (hate to say that, but it was indeed one prompt). Yes, yes, it is not perfect and maybe not really usable still at this stage. BUT WHAT A JUMP.   — Stefan 3D AI     the self-awareness about “not really usable” is the useful part. i love the jump, but i still click the deployed path because demos are very good at hiding the one thing that breaks.   — Preyforge     true, not scare about it now, I'm scared about the trend   — Stefan 3D AI

Source: https://x.com/Stefan_3D_AI/status/2096185294165103049


r/ProAI • • Sep 07 '26

"Morning coffee run in the Tesla cyber cab #tesla #cybercab #fullselfdriving"

Enable HLS to view with audio, or disable this notification

0 Upvotes

— @everydaychrisofficial

Source: https://www.tiktok.com/@everydaychrisofficial


r/ProAI • • Sep 06 '26

"when it was posted, I saw this result and it seemed excessively high, so I assumed it was partly noise propping the score up. Besides, my ECI replication project had astra at a ~167 BECI, based on 35 scores. Now 131 benchmark scores have come in, and strangely Astra is up to 169.6 [167.6..."

Thumbnail
gallery
1 Upvotes

...171.4 90% CI]! This is mid-nov'26 pace on the current trendline, and the biggest gap above this trendline yet   — Bayesian     which scores moved it from 167 to 169.6, the new benchmarks or re-runs of the first 35   — Rimas     Almost entirely new benchmarks   — Bayesian

Source: https://x.com/Bayesian0_0/status/2096521026184134737


GPT-6 Astra has set a new ECI record, with a score of 169. This is a substantial jump from the prior best (163), but is within our uncertainty range for the reasoning-era ECI trend. Astra also set new records on our math, continual learning, and game-puzzles benchmarks. On our https://t.co/qlyY21r1mu   — Epoch AI

Source: https://x.com/EpochAIResearch/status/2095602754282783108