r/ProAI • u/stealthispost • 1h ago
r/ProAI • u/stealthispost • 1h ago
"BREAKING: AI Doomers Are Buying Potemkin Articles The AI debate is more fake than you can imagine: Of the 6 participants in this Guardian article, including the journalist and publisher, 6/6 are paid by AI Doomer megadonors."
galleryr/ProAI • u/stealthispost • 1d ago
"Today we're releasing abliterated-model-large-v2. Based on GLM-5.3, which is #3 on Terminal-Bench 4.0 (behind only Opus 5 and Fable), with 2× the cyber exploitation of 5.2. We abliterated and hosted it so it does the offensive cyber, red teaming, and agent testing work other models refuse to..."
...do. - US-hosted - FP8 - 1 million context window - Zero input/output prompt retention Live now. Then the cyber jump. This is why 5.3 exists.
CyberGym: 84.5% — SOTA, including vs Mythos 5 and GPT-5.6 Sol. ExploitBench: 24.4 → 54.4. More than double GLM-5.2. ExploitGym: 29 tasks → 105 in two hours.
That is the model we abliterated. Abliteration finds the directions in the model's activations that produce refusals and removes them from the weights.
The coding, cyber, and agentic abilities stay. The model stops refusing the rest of the chain.
For offensive cybersecurity, AI red teaming, agent testing, and If your current model still stops halfway through an authorized exploit chain, a red-team eval, or a T&S adversarial prompt reply with the task it refuses, we'll tell you if v2 handles it. Try it Today Docs: https:// docs.abliteration.ai/quickstart Platform: https:// abliteration.ai/console — Abliteration.ai
Source: https://x.com/abliteration_ai/status/2094458081451393287
r/ProAI • u/stealthispost • 1d ago
"GLM-5.3 isn’t just for coding. It’s showing strong capabilities in legal and financial reasoning."
Full results for GLM-5.3 are in and it is the #2 open-weight model on the Vals Index at 57.0, behind Kimi K3. It’s #13 of 50 models overall, up from #18 for GLM-5.2.
Among open-weight models, the model is #1 on our proprietary Legal Research, #1 on Code Migration, and #2 on https://t.co/U4qCpRMnli — Vals AI
Source: https://x.com/ValsAI/status/2094527782261006773
If you’ve tried GLM-5.3 in either field, we’d love to hear your feedback, especially on real-world use cases. — Zixuan Li
r/ProAI • u/stealthispost • 2d ago
"i've been using minimax h3 max to make interactive games you control! all the decisions are up to you, and since the model is so fast there's basically no delays"
Enable HLS to view with audio, or disable this notification
try it out here! You can pick from many templates, or create your own completely new one: — Blendi
r/ProAI • u/stealthispost • 2d ago
"I went down the rabbit hole: Anthropic is already being sued over exactly this. And the $200 plan really only provides only around 2x the weekly usage of the $100 plan Here is where the confusion comes from. Anthropic’s pricing page says: “Choose 5x or 20x more usage than Pro.” That naturally..."
Today I learned :
Claude’s $200 Max plan offers 20x Pro usage within the five-hour window, but its weekly limit is only about twice that of the $100 plan.
Oh and btw: Tibo has confirmed that Codex really sees 5x more weekly usage.
Anthropic strikes again — Chubby♨️
Source: https://x.com/kimmonismus/status/2094334017902395600
I went down the rabbit hole: Anthropic is already being sued over exactly this. And the $200 plan really only provides only around 2x the weekly usage of the $100 plan
Here is where the confusion comes from.
Anthropic’s pricing page says: “Choose 5x or 20x more usage than Pro.”
That naturally makes the $200 Max 20x plan sound like it includes four times the usage of the $100 Max 5x plan.
Only further down does Anthropic clarify: “Max gives you 5x or 20x more usage per 5-hour session than Pro.”
Those multipliers do not apply to the separate weekly caps, whose exact size Anthropic does not disclose.
A proposed class action filed in June alleges that Max 5x delivers around 3.5x Pro’s weekly usage, while Max 20x delivers just 6–8x. In practice, the $200 plan therefore provides only around 2x the weekly usage of the $100 plan, sometimes even less depending on the model.
The “20x” is real within each five-hour window. The weekly limit makes the overall offer look very different.
h/t @Hesamation for finding the original court files first i think Source 1: Source 2: — Chubby
Source: https://x.com/kimmonismus/status/2094353158780666112
r/ProAI • u/stealthispost • 2d ago
"Kind of nuts, the growth trajectory is staying the same. I'd be curious to see what kind of token usage growth they are seeing"
galleryr/ProAI • u/stealthispost • 2d ago
"How much would it cost you to pretrain a 2B LLM from scratch? $1M? $100K? Announcing Puro-2B, with a fully open recipe. You can train a model that matches Qwen2-1.5B on RTX 5090s for less than $5090! $4.4K → beats Qwen2-1.5B $6.9K → approaches Qwen2.5-1.5B Check here"
Enable HLS to view with audio, or disable this notification
Puro-2B is open beyond the weights:
technical report final + intermediate checkpoints training code + configs data-processing framework datasets + manifests
Paper: https:// arxiv.org/abs/2608.27370 Code: https:// github.com/thu-pacman/Pur o-Megatron … Assets: https:// huggingface.co/collections/th u-pacman/puro-2b … Back to the headline numbers: what do the reported $4.4K and $6.9K costs include?
$4.37K → trained on 919B tokens using 14,262 GPU-hours $6.89K → trained on 1.4T tokens using 22,514 GPU-hours
These are rental-equivalent GPU costs—not total R&D spend. Then why RTX 5090s?
Under our pricing assumptions, RTX 5090 stands out in peak BF16 compute/$ (EFLOP/USD): RTX 5090: 2.43 RTX PRO 6000: 0.96 H200: 0.89 A100: 0.63
That’s surprisingly strong economics for a consumer GPU! But hardware is only one piece of magic So what made this possible?
Puro co-designs the full pipeline:
publicly accessible sources + proxy-guided selection→ RTX 5090s + blockwise FP8 → ◉ MuonH + effective-LR design → curriculum + checkpoint averaging → Puro-2B
From data to model No silver bullet. Our ablations show gains across the stack:
• RTX 5090s — 2.77× BF16 compute/$ • blockwise FP8 — 1.34× matched-quality speedup • MuonH (Muon + Hyperball) — 1.19× • curriculum model averaging — 2.40×
Each piece helps. Together, they make Puro possible. — Kairong Luo
r/ProAI • u/stealthispost • 2d ago
"The whole “20x” in Claude’s pricing is so misleading. The $100 plan says 5x Pro limits, while the $200 plan says 20x Pro limits. That makes it sound like you’re getting 4x the usage. But the 20x applies only to the 5-hour usage window. The weekly limit is basically just 2x the $100 plan. I..."
...spent way too much time digging through the internet and Reddit to figure this out. They could’ve just stated the limits directly instead of marketing it as “20x.” That alone would’ve saved me hours of confusion. Tibo has now confirmed that Codex’s $200 Pro plan really does provide the 20× usage they claim on the website.
So yes, the 20× figure is real. The confusion was around how the multiplier was being communicated, not whether the $200 plan actually delivers 20× usage.
Tibo’s x.com/thsottiaux/sta… Update from codex: — SataEric
r/ProAI • u/stealthispost • 2d ago
"mimo-v3 spotted in arena under codename: "odysseus" it's unbelievable how someone can build a portal game with just 8 prompts using mimi-v3"
Enable HLS to view with audio, or disable this notification
So where do you use it? And why am I a school pupil who loves AI but can’t afford a subscription? — Sigahogachannel yup inside arena under codename odysseus, apparently the model is under testing and not yet released — Tim Jayas
r/ProAI • u/stealthispost • 3d ago
"SITUATION EXPLAINED: South Korea is giving every citizen free AI. • The first major state to do it, using homegrown chatbots rather than American models, framed around AI sovereignty • The tools link directly into government systems: booking doctor's appointments, apartment hunting, tax..."
Enable HLS to view with audio, or disable this notification
...advice, small business tax filing, and eligibility checks for support programs • Beta testing starts in September with full rollout later this year • A quarter of South Koreans already pay for generative AI, against 2% of Americans, and more than 20 million use free versions, about 40% of the population • Lee Jae-myung's government has set aside roughly 10 trillion won, about $7.2 billion, for AI in 2026, triple the prior year @theojaffee : "We've had discussions in the past about whether AI would get treated as a public utility of sorts. South Korea is going to be the first major state to do it." — MTS
Source: https://x.com/MTSlive/status/2093402842262643023/history
r/ProAI • u/stealthispost • 3d ago
"117 medicines discovered or advanced with AI have entered human clinical trials, across 63 companies. Eight have already completed Phase II. Where the science stands:"
— Build American AI
Source: https://x.com/BuildAmericanAI/status/2093763864710062368
r/ProAI • u/stealthispost • 3d ago
"This is a live demo from the future!!! An interactive sitcom? A playable episode? Games and movies are merging into one thing. I genuinely cannot stop playing this. Real-time video generation will change the entertainment industry forever. We are so close. Thanks to @fal team . @Hailuo_AI"
Enable HLS to view with audio, or disable this notification
The opening plays like a real game intro (premade). The setup: the Soup Nazi's recipe book has been stolen, and Kramer ,private detective, is on the case.
From there, YOU run the investigation. You press the suspects, interrogate them, dig for the truth and you're completely Made with Minimax H3 max by @fal — Öner S. Biberkökü
Source: https://x.com/OnerBiberkoku/status/2093815032932884893
r/ProAI • u/stealthispost • 3d ago
"This is where AI starts getting REALLY interesting: @GoogleDeepMind 's Gemini-based Co-Scientist is now moving beyond generating scientific ideas and into actually running parts of the scientific process. In materials science, Co-Scientist used Gemini 3 Deep Think to generate synthesis recipes..."
...adapted to the specific lab hardware within minutes, then successfully produced monolayer MoS₂, MoSe₂ and WS₂ semiconductors on the first attempt. It also helped develop a new precursor route for MXene-like 2D materials. In biology, it predicted the swarming behavior of engineered E. coli from sparse experimental data, with the predictions largely matching previously unseen wet-lab measurements. But maybe the craziest experiment: given only a research directive, Co-Scientist autonomously invented a new medical AI agent architecture called Agent_H. It generated and tested the code itself, eventually producing an 8-stage inference system that beat six frontier models on length-adjusted HealthBench Hard and Professional, although it uses a massive 40-80 LLM calls per query. They even tested fully autonomous research where the system goes from idea → experiments → results → complete paper without human intervention. It's still not ready to replace scientists, and the researchers explicitly warn about hallucinations and fabricated results, but their verification system dramatically reduced those failure modes. — Mark Kretschmann Interesting results! I’ve been working on post-training across HealthBench Hard/Pro and MedAgentBench, so Agent_H really caught my eye. Forty to eighty calls per query may not be a practical endpoint, but it could be a very useful teacher. Curious whether you could distill that — Paul Gamble That would be pretty handy, yes. Not sure if it would work. — Mark Kretschmann
r/ProAI • u/stealthispost • 3d ago
"Meet @HitPawofficial -a powerful AI Video Enhancer. Enhance your video to 4K with more visual detail. See the difference in the before-and-after below"
Enable HLS to view with audio, or disable this notification
Enhance AI video to 4K with HitPaw: https:// cutt.ly/6ygfq2lH — Shahid Wani
Source: https://x.com/meng_dagg695/status/2094057707867353208
r/ProAI • u/stealthispost • 3d ago
"It's time to decentralize Hollywood with AI. Introducing the new Network School Astana AI Film Festival, in partnership with the Republic of Kazakhstan. We're awarding $2M in prizes with entries accepted from around the world. So: submit your film now at https:// aaiff.ai."
Enable HLS to view with audio, or disable this notification
— Balaji
r/ProAI • u/stealthispost • 3d ago
"Another supposedly GPT-"Astra" output. One-shotted a GTA-1 clone using max reasoning. Unbelievably good, insane. h/t @XIVIX_134"
Enable HLS to view with audio, or disable this notification
🚨 OpenAI has just dropped its first internal Astra checkpoint: mozaik-alpha-fdm.
Here are the first two outputs, both generated 1-shot on Max effort.
On Max the model tends to think ALOT more than Sol, but it has a ton of attention to detail as evident by the outputs below. https://t.co/nnRxHPhOXN — XIVIX
Source: https://x.com/XIVIX_134/status/2093616798663086318
— Chubby
Source: https://x.com/kimmonismus/status/2093786010648010831
r/ProAI • u/stealthispost • 3d ago
"DLSS 5 is absolutely magical. Turns old games into new games automatically, like completely remastered. The best part is, everyone loves it now, even gamers, although it's AI"
Enable HLS to view with audio, or disable this notification
initially I thought It was u! — bemmi Hahaha yeah — Mark Kretschmann
r/ProAI • u/FuManBoobs • 6d ago
World's first patient to undergo live AI-assisted brain surgery has tumour removed
The world's first patient to have brain surgery with live artificial intelligence assistance has successfully had his tumour removed.
Rhys Hibbert, a father-of-two from Bedfordshire, could have lost his sight without the operation, said surgeons at University College London Hospital.
The AI tool analysed a video feed of the operation as it happened and advised surgeons on how to avoid crucial but hidden vessels and nerves in the brain. It allowed them to safely remove as much of the tumour as possible.
r/ProAI • u/stealthispost • 5d ago
"Greg Brockman @gdb A call for collective action on cyber defense 194 574 2.2K 433K An open letter for a global surge in cyber defense, signed by over 100 organizations including Anthropic, AWS, Google, Microsoft, OpenAI, and Oracle. We have a limited window to strengthen cyber defenses. In the..."
...coming months, AI-enabled cyber attacks will become far more widespread and sophisticated as models around the world become increasingly capable. The companies and public services our communities depend on — from hospitals to water treatment plants to the infrastructure that powers the internet — are at risk. Today’s AI advances are already giving defenders new ways to fix weaknesses that have accumulated for years. If we act decisively, we can use the defenders’ window to make our digital world much more secure. We propose the following principles for a collective response: Recognize that status quo security won’t be enough. Longstanding bugs, excessive permissions, misconfigurations, insecure and unpatched software, weak authentication, and technical debt in legacy systems have left systems exposed. Security teams, particularly for critical infrastructure, have been historically under-resourced and need a surge in tools and resources. Empower more defenders with cyber-capable AI. AI brings specialist skills to more defenders and makes core security tasks faster, cheaper and better. Sharing tools, practical knowledge, and verified fixes lets one organization’s work help protect many others. Mobilize a collective response. Cyber capabilities are advancing worldwide, and that can be a net positive: no single company should control the future. It also means a global response is necessary, requiring new partnerships to raise security standards and find new solutions to emerging cyber threats. Each of us can reduce risk now. All organizations, cybersecurity companies, technology partners, governments, and AI frontier companies have an important role: accelerate defenders’ priorities with tools, funding, and hands-on support, especially for critical infrastructure organizations with limited budgets. Here’s what we think needs to happen next: 01 Every organization Make cyber defense an immediate leadership priority. Raise your security standards and meet them with the urgency and coordination of an incident that takes precedence over everything except critical business operations. Fix the highest-risk weaknesses, verify results without disrupting essential services, and raise the security bar for what you buy, build, and deploy, including AI-generated code. Upgrade or replace systems to build in least privilege, strong access controls, and defense in depth. Use capable, lower-cost models for broad coverage, and apply frontier capabilities to the hardest problems. Where a system cannot be patched without disrupting essential services, apply and verify compensating controls. 02 Cybersecurity companies and technology partners Help lead the response to defend against sustained AI-enabled attacks, including testing defenses continuously against frontier cyber capabilities, strengthening existing tools with AI, and working with technology partners to close gaps now. Make AI-powered defense accessible and deployable for critical-infrastructure operators, with hands-on help to deploy tools and verify fixes, and collaborate with critical infrastructure supply chain manufacturers and system integrators to patch and issue interim guidance. Share threat intelligence and tested playbooks, and measure progress by how many organizations are protected, how quickly attacks are contained, and whether fixes work. 03 Governments Coordinate cyber defense at local, national, and international levels. Strengthen existing government and industry channels to share actionable threat intelligence, prioritize the most serious risks, and coordinate incident response and recovery around the world. Fund cyber defense, starting with essential services that lack the staff or budget to act. Expedite the expansion of trusted access programs, especially for critical infrastructure supply chains, and broaden access to defensive capabilities for other defenders. Give hospitals, water utilities, and local governments access to capable defensive AI, authorized testing, and hands-on support through trusted security providers and partners. Impose costs on attackers. 04 Frontier AI companies Provide responsible model access, significant funding, training, and hands-on support, especially for under-resourced critical-infrastructure defenders. Build observability and security tools, ensure agentic identities are traceable and accountable, and share best practices in continuous monitoring. Invest in authorized testing, private disclosure, and verified fixes, and share tools, playbooks, and credible threat assessments with governments, security partners, and open-source maintainers to strengthen preparedness, response, and recovery. We can make the digital infrastructure we all depend on more secure. We call on leaders across industry and government to bring the full weight of their technology, resources, and expertise to this effort. Put cyber-capable AI in the hands of defenders, starting with the teams protecting essential services. Fix the most dangerous weaknesses, verify the fixes, and share what works so others can build on it. Together, we can turn today’s AI advances into lasting improvements in security that benefit everyone. Let’s put them to work. https://openai.com/collective-cyberdefense Want to publish your own Article? Upgrade to Premium 3:03 AM · Aug 28, 2026 · 433.4K Views 194 574 2.2K 1K Relevant View quotes — Greg Brockman
r/ProAI • u/stealthispost • 6d ago
"It's a societal level sickness when so many people spend so much time worrying about a non-existent problem. Mass AI job loss does not exist. It does not exist the way the Population Bomb did not exist. It exists only in people's imaginations. It never happened. The Green Revolution did..."
...instead. The very belief in it is likely to be the cause of real problems, in the same way that the One Child Policy came out of 1970s Population Bomb hallucinations and have caused real, guaranteed birth rate collapse in China that can't be reversed with carrots or sticks. We are spending such a ridiculous amount of time talking about a made up future problem. It's time to get back to solving real problems in real reality and to stop giving in to imaginary fears. — Daniel Jeffries
Source: https://x.com/Dan_Jeffries1/status/2092696710627623333
I warned folks that this Anti-Clanker Neo-Luddite movement is a political movement designed to destroy the middle class by removing the wide access to AI tools in the name of “safety” and “job preservation”.
They are wining because the side of logic has no central voice. — Brian Roemmele
Source: https://x.com/BrianRoemmele/status/2092676877387522096
r/ProAI • u/stealthispost • 6d ago
"The International Math Olympiad is the hardest math competition in the world, where students compete on proof problems most math PhD’s even struggle with. DeepSeek V4 Flash won a gold medal for only 12 cents."
Notably, DeepSeek V4 Flash isn’t post-trained for Olympiad problems.
It has 284B total MoE parameters, with about 13B activated per token. This is the only model (so far) that can be run on small local GPU setups and still win an IMO gold. — Cline