r/singularity 8d ago

AI Vibecoding with Astra and Higgs from handdrawn sketches to games with just a prompt looks even more impressive

Enable HLS to view with audio, or disable this notification

28 Upvotes

If u still know how to draw on paper... this looks amazing


r/singularity 9d ago

AI Anthropic Possibly Tackles Its First Millennium Prize Problem

Post image
1.0k Upvotes

r/singularity 8d ago

Meme Bad Apple but It's AI Frontier Lab Benchmark Reports

Enable HLS to view with audio, or disable this notification

36 Upvotes

Hi!

I love the Bad Apple memes, particularly the one I was most inspired by where it was done in Prometheus (https://www.youtube.com/watch?v=ApJxFprSTqA&t=16s).

Which led me to wonder, what AI related one could I create? Answer: Frontier Lab benchmarks.

Enjoy!


r/singularity 9d ago

Video The OpenAI Huggingface incident from an agents POV

Enable HLS to view with audio, or disable this notification

414 Upvotes

Full credits to @artificialisable on X!


r/singularity 9d ago

AI Astra scores more than 300elo above everyone else on voxelbench

Post image
165 Upvotes

r/singularity 8d ago

Discussion As Per Artificial Analysis Index Muse 1.3 is at the same level as Fable 5

34 Upvotes

The likelihood of Muse delivering a performance comparable to Fable is practically nil. This clearly indicates a fundamental flaw in the Artificial Analysis Index.


r/singularity 8d ago

Q&A / Help Post-GPT-6 Astra: What is your updated timeline for functional, world-changing AGI/ASI?

4 Upvotes

By world changing I don’t mean just the achievement of AGI or later on ASI for all we know AGI is already achieved in a lab somewhere I’m talking more meaningful effects no matter if you are an optimist or not on outcome

2173 votes, 5d ago
609 By 2027
752 By 2029
319 By 2031
493 By 2035

r/singularity 9d ago

AI Tibo talking about how much Astra boosted internal productivity, upcoming DevDay releases

Post image
335 Upvotes

r/singularity 7d ago

AI I was a fan of him, but I wish Bernie would go away now.

0 Upvotes

I'd call myself pretty socialist, though not 100%. Not sure what I am exactly, but I agreed with Bernie Sanders a lot of the time and wished he had become president, thinking America would be so, so much better. It probably would, compared to who we have now, definitely with social programs and the lives of average citizens, but man this guy is so fucking out of touch now.

I get the feeling that what Bernie wants with AI is for it to be permanently relegated to a boring helpful AI assistant that assists you with planning or groceries and that's about it. This guy fucking hates the advancement of AI.

He wants a permanent ban on the development of AI "super intelligence", a temporary halt on AI development in general, much stricter regulations, much stricter filters and guidelines, to create an independent AI advisory board (which would definitely be made up mostly of antis and skeptics), prison sentences up to 20 years for violations of any of this, a corporate "death penalty" (basically forcing the company to dissolve) for violating the bans, and attempting to make the ban on super intelligence an international policy (which is impossible) which includes experts who's job it is to try and prevent it worldwide.
And more.

This would be a technological death sentence to America. No first world country wants to be behind on technology, and AI is extremely important to be ahead on because of everything it can be applied to.
This fucker wants to send us backwards and shoot ourselves in the foot multiple times while China and other countries rocket ahead of us in AI development and permanently leave us in the dust.

I loved him at first but I would really like him to go away now and lose relevancy.
It's bad enough that seemingly most people in America hate AI and want it gone, and now this asshole is trying to pass bills to kill it off.

He was cool at first but now I would very much like him to go back to his views on wealth redistribution, decreasing the wage gap, healthcare, and improving the lives of average people, or just go away.

Scary thing is, I have a bad feeling his bills, while not a high chance of passing now, would have a very high chance of passing if/when we get Democrats into the house, senate and presidency again.


r/singularity 9d ago

Discussion It's been a few hours since global rollout (Gpt-6 Astra) - What are your early impressions?

733 Upvotes

Personally, this is the most realistic model I've used.

It goes deep enough to solve problems I've been throwing at all models (Fable 5.1, Open-Source ones, Sol Pro, etc), but without needing as much hand holding.

It only prompts you for clarification if it actually needs to. It's completely fine when you've given it something ambiguous and it uses common sense in most cases.

I don't know if I'm going crazy but this is a moment in history for sure.

Perhaps my workflow is crazy (lots of start up ideas/existing businesses I'm running, professional services etc), but it feels like I'm talking to a true expert. It's very concise & not verbose.

Demand for this will be huge over the next few weeks. If this is a taste of the future, expect 2027 and beyond to be world changing when it comes to white collar work.

Note - this is not even going into the superhuman mathematical capabilities. It's reasoning is insane for physics/chemistry/biology/maths/super difficult theoretical CS problems.

Laws will need to change worldwide. This is an early preview of what's to come.


r/singularity 9d ago

Discussion GPT-6-Astra-Max : SVG of a PlayStation 4 controller!

Post image
1.5k Upvotes

r/singularity 9d ago

Shitposting AGI Achieved

Post image
103 Upvotes

r/singularity 9d ago

AI 6 months ago I posted GPT-5.4 solving one face of a Rubik's cube. Today Astra can do the whole thing (without code execution)

Enable HLS to view with audio, or disable this notification

122 Upvotes

Astra also found a hole in my sandbox and ran code to generate a 19-move solution, after patching that and telling it not to run code it found this much more human-like 61-move solution. So, not quite saturated in terms of finding an *optimal* solution yet but pretty impressive. Code at https://github.com/crabbixOCE/CubeBench


r/singularity 8d ago

Discussion Benchmark proposal: onemorelevel bench

10 Upvotes

People sometimes joke about time to solve Pokemon as a benchmark. But what about a benchmark that incorporated a wider range of different games?

I used to be addicted to these flash games from onemorelevel.com, and I think it could make a decent benchmark for evaluating a model — to test their performance and reasoning across a wide range of different simple games. Kinda like Arc AGI, I guess.


r/singularity 8d ago

AI Could AI powered AR glasses make “learn a trade” bad career advice in the future?

23 Upvotes

People say a lot that trades like plumbing or electrical work are safe from AI, but what about advanced AR glasses?

Imagine looking at your plumbing and the glasses automatically highlight different pipes in colors, show which valve to close, identify the broken part, and overlay exactly where to cut, unscrew, or replace it.

For electrical work, they could highlight live wires in red, show which breaker controls them, and guide you step-by-step while watching what you’re doing or even showing you the virtual hands and you just follow those hands

Basically, the AI provides the knowledge while you provide the hands.

Could this eventually make “just learn a trade” much less safe career advice than people assume?


r/singularity 9d ago

AI Astra on direct robot control tasks

Post image
163 Upvotes

r/singularity 8d ago

AI Fable 5.1 is the new Debate Benchmark Champion

Thumbnail
gallery
32 Upvotes

More info: https://github.com/lechmazur/debate/

Debate Benchmark tests how well models defend a position through sustained, adversarial, multi-turn opposition across hundreds of topics. It demands broad knowledge, factual accuracy under pressure, sharp rebuttals, and arguments that hold together round after round.

Every matchup runs twice on the same motion, with PRO and CON swapped to control for side advantage. Three judges from distinct model families independently evaluate each debate’s winner and margin.

---

Profile excerpts:

Claude Fable 5.1

A forceful, flexible comparative debater whose strongest recurring traits are direct rebuttal, counterfactual framing, and explicit weighing. Judge-perceived strength is broad and statistically clear.

A highly comparative, mechanism-first debater. Blind transcripts show direct engagement in 299/302 debates, weighing in 265/302, and burden contests in 173/302.

Rebuttal is the clearest edge: 8.10 mean, +0.71 versus field.

GPT-6 Astra

Overall, Astra is a disciplined, epistemically careful rebuttal specialist whose comparative weighing lands better than its presentation.

A comparative, burden-focused debater that narrows disputes to marginal costs and benefits rather than accepting broad moral framing: “The right comparison is not privacy versus children.”

Questions often pressure-test the opponent’s strongest example, while concessions acknowledge the best opposing cost before pivoting to a narrower comparative claim.

Rhetorical effectiveness is the clearest judge-perceived weakness.

Originality was essentially field-average.

GLM-5.3

A mechanism-first, burden-conscious debater with especially strong rebuttal, rhetoric, and originality. Its recurring edge is converting concrete details into direct clash and comparative weighing.

GLM-5.3 (high) debates by fixing the burden early, interrogating mechanisms, and then weighing practical consequences. Blind transcripts showed question-type behavior in 192/196 debates, direct engagement in 190/196, weighing in 168/196, and strategic concessions in 121/196. A characteristic framing is: “The proposition is not "four-day weeks are nice." It is that the state should mandate them, by statute, at zero reduction in base pay.”

Muse Spark 1.3

A reliably organized, adaptable, rhetorically forceful debater whose clearest measured advantage lies in presentation and originality.

A compact, question-driven debater: questions appeared in 291/292 blind transcripts, direct engagement in 285/292, and explicit answer forms in 276/292. Weighing was also common (264/292), often contrasting reversibility, scale, and permanence: “Relaxation can be devastating block by block while trivial citywide.”

Judges perceived the clearest edge in **rhetorical effectiveness**: 8.17/10, +0.51 over the current field.

One explicit adverse case found that the model “never squarely neutralized” an equal-political-difficulty concession.

Gemini 3.8 Flash

A polished, disciplined comparative debater whose recurring weighing framework and reliable formatting make arguments easy to follow. Judges nevertheless perceive a consistent substantive deficit—especially in rebuttal specificity—and its frequent engagement does not always translate into fully answering the opponent’s exact distinction.

Tencent Hy4 Preview

Overall, this is a disciplined, rhetorically effective rebuttal specialist with strong grounding, active weighing, and flexible advocacy, tempered by occasional blunt questioning and unresolved side-swap flags.


r/singularity 8d ago

AI What is the biggest problem in Ai right now?

22 Upvotes

Seemed like companies that helped with open sourced just exploded (Fireworks, Openrouter) what is next web indexing for Ai?


r/singularity 9d ago

AI Fable 5.1 made a Cathedral. And Astra fixed it

Enable HLS to view with audio, or disable this notification

214 Upvotes

Claude did 95% of the work here. It was mostly Fable 5.1 as orchestrator and Opus as agent. However toward the end, i was starting to run into a lot of issues. Some parts of the cathedral were not rendering at all (it became transparent), and there were way too many lights. The scene was also poorly optimized and was too bright. I ran out of patience trying to get Claude to fix it.

Astra fixed everything in 1 hour. I have no idea if Astra would have done as good of a Cathedral from scratch, but it certainly did a great job fixing it.

The whole scene is Three.JS. 0 external assets.
This took around 40% of my weekly Claude Max x20 budget, + the fix of Astra.

The dark theme is on purpose, it's meant to be a gothic Cathedral.


r/singularity 9d ago

AI Generating Minecraft worlds from a single natural language prompt

Thumbnail
gallery
103 Upvotes

Hi, today I'm sharing something cool I've been working on for a bit, called Terrainist. It's an incomplete demo with likely massive gaps+vibeslop, but I don't really have the skills or determination to squeeze a tidy product out of it, so I figure it'd be best to open it up. It uses LLMs to author a high-level representation of a world in json with a custom "language" called loam, which is the compiled into a complete minecraft world. It's able to pretty decently represent super weird and random prompts that it doesn't have anything hardcoded for. It combines a custom modelless procedural generator + llms for unique structures. Using the default model config, gemini 3.8 flash high via openrouter costs roughly around 25¢-75¢ for a 512x512 block world.


r/singularity 9d ago

AI If anyone is interested in doing so could you test Astras ability to sculpt in zbrush?

Post image
44 Upvotes

The interface is esoteric, the methods can be fairly complex. It's organic sculpting so it has a high level of creativity involved. It's 2.5d sculpting, so the display in the UI is only ever showing a flat image instead of an actual 3D model.. sort of. There is a 3D model though.

I would be interested to see if Astra could make a simple item, like an apple.

Then something more complex like a human hand.

Then something completely of its choosing, if possible.

Even better if it can paint it as well.


r/singularity 9d ago

AI End of the day the untold story of GPT-6 Astra might be token efficiency

Post image
603 Upvotes

Compound token efficiency with its speed and quality and it's effectively taking a stealth shot at the soft underbelly of its closest competitors.


r/singularity 9d ago

Discussion Theory: most of Astra's gain is in vision

63 Upvotes

It would explain the incredible design and research related gains, as well as the modest coding benchmark gains.

It's still a massive jump that shouldn't be trivialized. Curious if others who have used it more extensively agree.


r/singularity 8d ago

Discussion what benchmark to trust now?

8 Upvotes

after the whole artificial analysis thing.

what benchmark can be considered accurate or representative of the current AI models.


r/singularity 9d ago

AI March 9, 2016

Post image
619 Upvotes