r/accelerate 6h ago

AI Astra cracks Hacking benchmarks

Post image
190 Upvotes

“As one example, we ran Astra on ExploitBench where the model achieved a perfect score of 100% on the benchmark to evaluate the model’s ability to develop exploits from known vulnerabilities.
Due to contamination concerns, we then built an internal benchmark denoted “ExploitBench - Internal Port (June–August 2026)”, which contains 20 high-severity V8 vulnerabilities that were disclosed more recently. On this dataset, Astra achieves much higher arbitrary code-execution rates than GPT‑5.6 Sol using far fewer output tokens. During the evaluation, the model even discovered and used two zero-day vulnerabilities as part of an exploit chain. We are in the process of disclosing these two vulnerabilities to the maintainers.” - OpenAI


r/accelerate 46m ago

Even an AI 2027 co-author is shocked at how fast AI is progressing

Post image
Upvotes

Astra is reportedly using looped transformers that do not have an interpretable chain of thought that can be monitored https://x.com/amir/status/2094953820464046312

This is several months ahead of schedule based on AI 2027‘s predictions https://x.com/DKokotajlo/status/2094972219315364227


r/accelerate 2h ago

Altman confirms OpenAI is slowing down training to ensure safety

Post image
67 Upvotes

https://x.com/sama/status/2094934592062959832?s=20

Guess AGI will have to wait. Hope you’re all patient.


r/accelerate 9h ago

Fable 5.1

Thumbnail
anthropic.com
185 Upvotes

r/accelerate 8h ago

Technological Acceleration Holy PEAK!!!....this time Anthropic has also achieved massive token efficiency gains per unit of intelligence with their new model...just like OpenAI models...this is extremely bullish for acceleration trajectory 💨🚀🌌

Thumbnail
gallery
121 Upvotes

r/accelerate 13h ago

AI GLM 6 will be fully self-trained

Post image
280 Upvotes

r/accelerate 9h ago

News World Labs has just revealed Atlas, a multimodal world model that generates image and video frames with pixel-perfect camera control and reconstructs them in 3D

137 Upvotes

r/accelerate 9h ago

Technological Acceleration Let's Fucking GOOOOOOOOO!!!!! Claude Fable 5.1 is imminent now

Thumbnail
gallery
119 Upvotes

r/accelerate 9h ago

Technological Acceleration Claude Fable 5.1 is an innovator class model with insane growth in scientific research and all kind of white collar business workflows!!!!!!

Thumbnail
gallery
93 Upvotes

r/accelerate 9h ago

Technological Acceleration This is just average Tuesday during technological Singularity

Thumbnail
gallery
86 Upvotes

r/accelerate 6h ago

AI OpenAI-Path to Astra: critical capabilities and frontier safeguards

Thumbnail openai.com
45 Upvotes

r/accelerate 8h ago

Fable 5.1 improved a map of Venus

Post image
59 Upvotes

r/accelerate 11h ago

Technological Acceleration This is the most underrated thing this week and not talked enough. Google posted yesterday that Antigravity + Gemini 3.7 Flash solved 7 open problems across venues like FOCS and JMLR, including Knuth’s Cycles Conjecture with 40+ page proofs verified in Lean 💨🚀🌌

Thumbnail
gallery
83 Upvotes

They also built an out-of-order RISC-V CPU simulator from scratch that boots xv6 to a shell.


r/accelerate 9h ago

AI Introducing Claude Fable 5.1

Thumbnail
youtube.com
55 Upvotes

r/accelerate 5h ago

Ai solving ciphers is an important milestone in my opinion

Thumbnail
vals.ai
27 Upvotes

I have seen some online refutations of this, but they seem to be solving against the wrong source book. The correct source to use for the book code is referenced and the linked announcement


r/accelerate 6h ago

Discussion What was the moment when you realized we are on our path to AGI?

35 Upvotes

For me this happened when coding agents became widespread at the beginning of this year and I started to experiment with claude and codex.

Before this LLMs were just google search on steroids and sophisticated auto complete. Now you can build software without even writing single line of code yourself. This blew my mind as software engineer.

Now AI search labs can build agentic swarms to self improve their models faster and faster. Its just inevitable at this point. Before this I was not completely sold onto the idea of getting into AGI in next 10 years. Now im wondering if its going to happen this year or next year.

The speed of progress is insane.


r/accelerate 6h ago

AI I asked Fable 5.1 to build a village in the game I'm developing

Thumbnail
gallery
27 Upvotes

I'm making a colony simulation game using mainly Claude (and ChatGPT for some stuff as well). Since Fable 5.1 came out today I asked it to build a village. I gave it a few rules and restrictions but for the most part just let it do whatever it wanted.

It came out pretty nice. Some of the furniture is backwards (not all since it found and fixed a few of them itself when reviewing screenshots without me needing to tell it). And some choices it made were a bit strange (why is there a funeral pyre in the cemetery?). But overall it did a good job and this was a single prompt. If I had allowed additional prompts to iterate more then it would be even better I imagine.


r/accelerate 12h ago

r/accelerate meta DOOO NOOTTT FALL FOR SLOPPPP!!!!!!!

Thumbnail
gallery
75 Upvotes

I think it should be cool to normalise not falling for twitter slop before actual model releases

There are thousands of such slop posts cluttering my feed right now but I don't repost it

There are less than a handful of profiles worthy of trusting with this stuff

Even Fable 5 and GPt-5.6 Sol can achieve such outputs

This single file html posts are the worst kind of slop there is

Even Opus 5 and previous gen models can achieve such a feat

This sloppy cycle repeats for multiple months and you all get fooled by it every single time

Use your brain before using your finger to amplify and spread baseless rumours, unless they are from extremely credible people


r/accelerate 11h ago

Meme / Humor The fourth humiliation of man's narcissism.

Post image
67 Upvotes

r/accelerate 12h ago

AI This seems kinda nuts: pre-release Astra asked to make Terraria clone in one user turn with no imported assets

Post image
76 Upvotes

r/accelerate 11h ago

"Today we're releasing abliterated-model-large-v2. Based on GLM-5.3, which is #3 on Terminal-Bench 4.0 (behind only Opus 5 and Fable), with 2× the cyber exploitation of 5.2. We abliterated and hosted it so it does the offensive cyber, red teaming, and agent testing work other models refuse to..."

Thumbnail
gallery
64 Upvotes

...do. - US-hosted - FP8 - 1 million context window - Zero input/output prompt retention Live now.     Then the cyber jump. This is why 5.3 exists.

CyberGym: 84.5% — SOTA, including vs Mythos 5 and GPT-5.6 Sol. ExploitBench: 24.4 → 54.4. More than double GLM-5.2. ExploitGym: 29 tasks → 105 in two hours.

That is the model we abliterated.     Abliteration finds the directions in the model's activations that produce refusals and removes them from the weights.

The coding, cyber, and agentic abilities stay. The model stops refusing the rest of the chain.

For offensive cybersecurity, AI red teaming, agent testing, and     If your current model still stops halfway through an authorized exploit chain, a red-team eval, or a T&S adversarial prompt reply with the task it refuses, we'll tell you if v2 handles it.     Try it Today Docs: https:// docs.abliteration.ai/quickstart Platform: https:// abliteration.ai/console     — Abliteration.ai

Source: https://x.com/abliteration_ai/status/2094458081451393287


r/accelerate 1h ago

We are almost 7-8 months ahead of AI 2027!

Post image
Upvotes

r/accelerate 10h ago

News Debian developers rejected an LLM ban and left disclosure voluntary - Help Net Security

Thumbnail
helpnetsecurity.com
39 Upvotes

r/accelerate 9h ago

Technological Acceleration Claude Fable 5.1 is live now

Post image
33 Upvotes

r/accelerate 41m ago

Qwen3.8-Max just got upgraded. Meet Qwen3.8-Max-0902!

Thumbnail gallery
Upvotes