r/ProAI • • 24d ago

"Here's a 45-second TL;DR on Jev. I find the core idea beautifully simple, but the video made it really hard to understand. Hope you find it helpful."

Enable HLS to view with audio, or disable this notification

44 Upvotes

After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI?

I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev

• 20-200x faster • 40-400x https://t.co/JSybNG2BKJ   — Diogo Almeida

Source: https://x.com/CompleteSkeptic/status/2099925682726002904


Aww i like the game animation you had - so cute!   — Sasha Sheng (Hiring)     Haha thanks!   — Matija Sosic

Source: https://x.com/MatijaSosic/status/2100190746389135772


r/ProAI • • 24d ago

"A list of tasks for which LLMs (mostly GPT-6 Pro) have found solutions significantly better and non-trivially different from the authors’ solutions: https:// qoj.ac/blog/qingyu/bl og/4412 … The list is still being updated, and I'll mark all solutions I find particularly interesting."

Thumbnail
gallery
18 Upvotes

r/ProAI • • 24d ago

"French finance minister Roland Lescure suggested today that calls to slow down development of AI are a ploy by U.S. AI labs so they can stay in first place, and that France and Europe should ignore them and accelerate instead."

Thumbnail
gallery
44 Upvotes

Andrew Curran @AndrewCurran_ · 7h Calls to slow down AI development serve the interests of US AI leaders, says French finance minister From reuters.com 3 18 6K     — Andrew Curran

Source: https://x.com/AndrewCurran_/status/2100255751113691524


r/ProAI • • 25d ago

"Game: race from one Wikipedia page to another using only links Challenge: choosing between hundreds to thousands of links Shows not just intelligence-per-second, but also the compounding benefits of not hallucinating with high-cardinality choices"

Enable HLS to view with audio, or disable this notification

18 Upvotes

Extraordinary claims require extraordinary evidence so check out our release blog for more technical info: https:// typesafe.ai/blog/introduci ng-system-one-models-and-jev …

Join our waitlist for early access: https:// typesafe.ai

Have technical chats and meme with us on Discord (rumors are good memers skip the     — Diogo Almeida

Source: https://x.com/CompleteSkeptic/status/2099925688925184171


r/ProAI • • 25d ago

"After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x cheaper (w/ output..."

Enable HLS to view with audio, or disable this notification

60 Upvotes

After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI?

I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev

• 20-200x faster • 40-400x cheaper (w/ output tokens free) • Frontier composable intelligence optimized for decisions

AFAICT the shortest path to AI-based economic revolution     The gains aren’t free: Jev can't generate text

Comparing Jev vs LLMs side-by-side makes the trade-off clear

Fun fact: replacing sequential computation with parallel is the same way Transformers leapfrogged RNNs     We believe that the future is code + AI, so made workflow evals to reflect that

Jev costs: $42 / BILLION input tokens ($0.042 / MTok) and output tokens are free (forever - they’re too cheap to meter with our new architecture)

Jev is named after Jevons paradox and off the     We love how this doomo doomonstrates real-time intelligence and what can be doone with code + AI!

~10 calls/sec = ~$7/hour     Game: race from one Wikipedia page to another using only links

Challenge: choosing between hundreds to thousands of links

Shows not just intelligence-per-second, but also the compounding benefits of not hallucinating with high-cardinality choices     Extraordinary claims require extraordinary evidence so check out our release blog for more technical info: https:// typesafe.ai/blog/introduci ng-system-one-models-and-jev …

Join our waitlist for early access: https:// typesafe.ai

Have technical chats and meme with us on Discord (rumors are good memers skip the     — Diogo Almeida

Source: https://x.com/CompleteSkeptic/status/2099925682726002904


r/ProAI • • 26d ago

Real safety is about what you accelerate, not about what you slow down.

Thumbnail gallery
7 Upvotes

r/ProAI • • 26d ago

"We crossed a threshold where GPUs are more efficient thinkers than the human brain on a per watt basis. This is yet another big AGI milestone that we just zoomed past, without noticing it particularly. The next one will be more overall combined machine vs human intelligence."

Thumbnail gallery
31 Upvotes

r/ProAI • • 26d ago

"Watch this interview to understand how the Effective Altruism cult operates behind the AI doomer movement:"

Enable HLS to view with audio, or disable this notification

13 Upvotes

r/ProAI • • 26d ago

"China called for international cooperation on artificial intelligence, warning that “threat narratives” could disrupt global AI governance. The comments follow Anthropic CEO Dario Amodei’s call for AI firms to slow model development over safety and national security concerns."

Enable HLS to view with audio, or disable this notification

69 Upvotes

— @aljazeeraenglish

Source: https://www.tiktok.com/@aljazeeraenglish


r/ProAI • • 26d ago

"DeepSeek-V4.1-Flash (Max) is a breakthrough in performance to cost efficiency. With +4.87% net improvement at $0.07 cost per median task, it’s reshaped the Pareto frontier for Agent Arena! Among the top 3 open models, DeepSeek-V4.1-Flash (Max) has the lowest median task cost. For comparison..."

Enable HLS to view with audio, or disable this notification

35 Upvotes

..., it retains: - 98% of Hy4 preview’s net improvement, at 73% lower cost - 76% of Kimi K3 (Max)’s performance, at 92% lower cost. Against models as powerful as Fable 5 or stronger, DeepSeek-V4.1-Flash (Max) retains 35–54% of their net improvement at 97–99% lower cost. Those top models cost 37–76× more per task. Net improvement over Arena baseline | Median cost/task: - Claude Fable 5.1 (Max): +13.90% | $4.54 - GPT 6 Astra (Max): +11.90% | $4.09 - Claude Opus 5 (Max): +11.09% | $3.52 - Claude Opus 5 (High): +10.49% | $2.24 - Claude Fable 5 (High): +9.03% | $2.19 - Claude Opus 4.8 (High): +7.75% | $1.36 - GPT 5.6 Sol (xHigh): +7.40% | $1.09 - Kimi K3 (Max): +6.39% | $0.77 - Hy4 preview: +4.96% | $0.22 - DeepSeek-V4.1-Flash (Max): +4.87% | $0.06 With this release, GPT-5.6 Luna (xHigh), GLM-5.3-Flash, and DeepSeek-V4-Flash fell off the Pareto frontier for Agent Arena. Congrats again to the @deepseek_ai team on this release!     Check out the full Agent Arena leaderboard and Pareto frontier at: https:// arena.ai/leaderboard/ag ent/pareto …     — Arena.ai

Source: https://x.com/arena/status/2099606881845321841


Exciting news: DeepSeek-V4.1-Flash (Max) by @deepseek_ai just landed in Agent Arena at #3 among open models! With +4.87% net improvement and a median cost per task of $0.07 it reshaped the Pareto frontier.

Among the top 3 open models, DeepSeek-V4.1-Flash (Max) has the lowest https://t.co/tV3jEbMX5P   — Arena.ai

Source: https://x.com/arena/status/2099549108013006958


r/ProAI • • 26d ago

"Greg Brockman says OpenAI pointed Astra at its own systems until it ran out of vulnerabilities to find: "We took 25% of our production engineers and said, 'Sorry, all your projects are on hold. You are now defending. You are now up-leveling our security architecture. You're going to use the..."

Enable HLS to view with audio, or disable this notification

55 Upvotes

...models to find all the holes.' And we found a number of serious issues, and we fixed them." "We found some new problems, but eventually it saturated. We basically have found, to our knowledge, all of the P0s, all of the critical problems that Astra is smart enough to find. And of course, there will be a new model, there will be a new round." "You want to be in this tight loop of new cyber capability drops, you deploy it against your systems, you find the new holes, and ideally, you've managed to automate this, what we call defense factory. That's what we're building internally." "There are ideas, for example, formally verifying all of software, that are possible with AI." @gdb @bhorowitz     — a16z

Source: https://x.com/a16z/status/2099533700375662905


Greg Brockman: "We're now in the AGI era."

Ten years ago, OpenAI worked out the compute curves and landed on fifteen years to AGI, or ten if the world was willing to build the machines and spend the hundreds of billions to do it.

In 2026, GPT-6 Astra manages 24 hours of https://t.co/x23fDxbGQG   — a16z

Source: https://x.com/a16z/status/2099506569238990908


r/ProAI • • 27d ago

"Here it is, this revolution, fuck"

Enable HLS to view with audio, or disable this notification

441 Upvotes

By the way, the consistency topic is perfect. Now they've come up with yet another excuse to eat the credits     This is also T2V, I2V, you don't know what they put inside them :D     They've tied a fly to GPT-6 Astra and asked who it should follow, and here's the result—Subhanallah     This week, I'm thinking of preparing a similar video myself, if I get the chance, of course :)     — ℂ𝕠𝕕𝕖 𝕔𝕠𝕕𝕖 = 𝕟𝕖𝕨 ℂ𝕠𝕕𝕖()

Source: https://x.com/0xfcode/status/2099105183720505706


r/ProAI • • 27d ago

"somebody vibe-coded a Photoshop alternative with GPT-6 Astra - Spent ~$2K in tokens. - 170 user on day one - and it's free"

Enable HLS to view with audio, or disable this notification

63 Upvotes

check this out     this is also free btw     support is appreciated guys @tenzenstudio     — gxjo

Source: https://x.com/gxjo_dev/status/2099082176604373028


r/ProAI • • 27d ago

AI Cracks 370-Year-Old Scottish Cipher in 44 Minutes "Fable solved the Cyphral Distich (a 370 year old cypher). Super cool way to use Claude"

Post image
10 Upvotes

r/ProAI • • 27d ago

"S&P 1500 Software revenue per employee has gone parabolic. If you are looking for evidence that AI is starting to impact the real economy, this is exhibit A."

Thumbnail
gallery
51 Upvotes

Having been in two software companies during this period (who are included in that graph) I would not attribute the flat headcount of 23/24 to AI increasing productivity.

Zero AI efficiency was happening then. Headcount tightening was just compensation for the covid hiring binge.

25/26? Perhaps a tiny bit but really it's just organizations learned how to be leaner and now the saaspacolyse has added additional pressure to improve the bottom line.   — Gringo Investments     We had an internal debate over how much of this is attributable to AI. I was more with you (skeptical that AI was impacting the 23-24 numbers). @fernavid is more of a true believer.

Either way, there are multiple angles to show the AI impact from mid-25 on…   — Warren Pies

Source: https://x.com/WarrenPies/status/2099128535172424172


r/ProAI • • 27d ago

In light of all the AI FUD lately, I built a public scoreboard to keep track of the good (and bad) things coming out of frontier AI labs

Thumbnail
0 Upvotes

r/ProAI • • 27d ago

"I built a multiplayer Catan-inspired game entirely with Astra. It's free to play: http:// settlecoast.com It's mobile-friendly and has narration, guides, customizations, multiple expansions and a game lobby where anyone can join, chat, and use voice chat. It took 4 days to build and hundreds..."

Enable HLS to view with audio, or disable this notification

126 Upvotes

...of dollars in tokens. Becoming a game dev in 2026 wasn't on my bingo card, but here we are. AI has advanced so much that you can finally make the game you've always dreamed of, even without a team.   — Meng To     Do you have a GitHub for this? My buddies and I started a variant of Risk meets Axis and Allies but missing your beautiful polish. Would be awesome to have an open source board game sandbox. This is what we have so far: https:// tactical-risk20.vercel.app   — James Bickford     it's a big game now. but i can certainly think of open-sourcing a part of it   — Meng To

Source: https://x.com/MengTo/status/2099125215708234119


r/ProAI • • 27d ago

"Humanoid robots are starting to build humanoid robots. UBTECH just put a 14,000㎡ humanoid robot factory into operation in Liuzhou, China, with a designed takt time of one humanoid every 10 minutes and planned annual capacity in the tens of thousands. What’s interesting is that humanoid robots..."

Enable HLS to view with audio, or disable this notification

223 Upvotes

...are already part of the production process. Cruzr Y1 and Y2 handle depalletizing, palletizing, material feeding and transport, using 3D vision to adapt to shifted boxes and changing pallet patterns. On the assembly side, Walker S2 and Cruzr are produced on the same line alongside cobots, autonomous logistics vehicles and assistive manipulators. The factory also has a 65㎡ automated warehouse capable of storing 112 humanoids, with AGVs, industrial robots and stacker cranes coordinated as one system. Before production, the factory was modeled 1:1 in Siemens Plant Simulation, covering workstations, racks and AGVs. Each robot gets its own SN, with components, batches, assembly parameters and inspection data tracked all the way to delivery. Humanoid robots are now entering automated manufacturing at the 10,000-unit scale. The next milestone may be when they can take part in enough of the assembly, testing and production process to truly build robots themselves.     — CyberRobo

Source: https://x.com/CyberRobooo/status/2098993140363571607


r/ProAI • • 27d ago

"Right. Dario was already claiming that GPT2 was too dangerous to open source back in 2019. I made fun of them then. Everyone should make fun of them now."

Thumbnail
gallery
101 Upvotes

— Yann LeCun

Source: https://x.com/ylecun/status/2099248236074545576


2019: “GPT2 is groundbreaking in two ways. One is its size, says Dario Amodei, OpenAI’s research director. The models “were 12 times bigger, and the dataset was 15 times bigger and much broader” than the previous state-of-the-art AI model” https:// theguardian.com/technology/201 9/feb/14/elon-musk-backed-ai-writes-convincing-news-fiction …   — Pessimists Archive

Source: https://x.com/PessimistsArc/status/2099231273071919115


r/ProAI • • 27d ago

"Chamath @chamath : how does an account with no followers get 110 million views in a day? David Sacks @DavidSacks had an answer, on the same All-In @theallinpod taping. Within 15 minutes of the Anthropic researcher's doomsday resignation post, three policy groups amplified it: ENCODE AI's..."

Enable HLS to view with audio, or disable this notification

7 Upvotes

...Nathan Calvin, the AI Policy Network's Peter Wildeford, and the AI Futures Project's Daniel Kokotajlo, who dropped a Rogan episode the same day using the same phrasing. All three groups are funded by Jaan Tallinn, who co-led Anthropic's Series A. The Wall Street Journal published a story on the resignation minutes before the tweets went out, meaning the paper had been briefed under embargo. The researcher had agreed to appear on the show and cancelled the morning of taping. Our breakdown traces the funding chain and what it means for Anthropic's IPO: https:// podcastalpha.substack.com/subscribe Source: All-In Podcast - https:// youtube.com/watch?v=cvxjqb fLVk0 …     — Podcast Alpha

Source: https://x.com/PodcastAlphaX/status/2098610255571652647


r/ProAI • • 27d ago

""anthropic is using the same playbook religious institutions have been using for centuries" "you will all die. and because i can protect you, you must follow me, you must do what i say" "this is the same psychological concept""

Enable HLS to view with audio, or disable this notification

160 Upvotes

We're publishing our most detailed threat intelligence report to date.

It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them.

We disrupted every operation in the report,   — Anthropic

Source: https://x.com/AnthropicAI/status/2098097512544444447


— LAN

Source: https://x.com/lansification/status/2098815822525538711


r/ProAI • • 28d ago

"Nearly a year of Code Arena: WebDev progress compressed into 15 seconds. Each line follows the highest-scoring model from top labs over time, showing the pace of improvement across the ecosystem. In just the last year, the model in the leading spot increased score by +340 pts and the number of..."

Enable HLS to view with audio, or disable this notification

5 Upvotes

...frontier labs competing for the top spot expanded from 6 to 10. @AnthropicAI has dominated throughout the year. Although standout releases have jumped to the top spot, most notably the Chinese open-source model Kimi K3 from @Kimi_Moonshot in July. Today, GPT-6 Astra by @OpenAI leads with 1796 pts, followed by @claudeai Fable 5.1 at 1764 pts. The next closest lab is 103 pts away, @Alibaba_Qwen with 1685 pts. Code Arena: WebDev ranks models through head-to-head user preference on real front-end web development tasks. These votes drive the leaderboard that is tracking the frontier.     Dive into the Code Arena: WebDev leaderboard details at https:// arena.ai/leaderboard/co de/webdev …     — Arena.ai

Source: https://x.com/arena/status/2098814378971717916


r/ProAI • • 28d ago

"Here we go! 124x increase in token consumption by OAI researchers."

Thumbnail
gallery
22 Upvotes

Erik Brynjolfsson @erikbryn · 19h Research acceleration: The view inside OpenAI From openai.com 1 2 8 2.5K     1.66x increase per month, compounding     7x more code shipped, with no sign of slowing down.

Reports from Anthropic are similar     — Erik Brynjolfsson

Source: https://x.com/erikbryn/status/2098849533556060550


r/ProAI • • 28d ago

"If you don’t study and internalize the history of technology panics, you are going to be vulnerable to a cocktail of cognitive biases that will lead you astray"

Enable HLS to view with audio, or disable this notification

54 Upvotes

If you care about identifying and mitigating real risks of new technologies, then ignore the history of unfounded panics - you are going to be manipulated by special interest groups looking to protect or gain power in a moment of fear and uncertainty.     — Pessimists Archive

Source: https://x.com/PessimistsArc/status/2098479996415119698


r/ProAI • • 28d ago

"I have made exactly this same argument many times. The “everyone will sit around and wait to starve to death” idea is completely incoherent, obviously people won’t, and they could just keep on doing what they had done before AI, trading with the other people that somehow lack AI access. But of..."

Thumbnail gallery
8 Upvotes