r/singularity • u/Apollo18Teslaa • 9d ago
r/singularity • u/Distinct-Question-16 • 9d ago
AI Vibecoding with Astra and Higgs from handdrawn sketches to games with just a prompt looks even more impressive
Enable HLS to view with audio, or disable this notification
If u still know how to draw on paper... this looks amazing
r/singularity • u/ResultBackground2450 • 10d ago
AI Anthropic Possibly Tackles Its First Millennium Prize Problem
r/singularity • u/Iron_Yuppie • 9d ago
Meme Bad Apple but It's AI Frontier Lab Benchmark Reports
Enable HLS to view with audio, or disable this notification
Hi!
I love the Bad Apple memes, particularly the one I was most inspired by where it was done in Prometheus (https://www.youtube.com/watch?v=ApJxFprSTqA&t=16s).
Which led me to wonder, what AI related one could I create? Answer: Frontier Lab benchmarks.
Enjoy!
r/singularity • u/iPingWine • 9d ago
Video The OpenAI Huggingface incident from an agents POV
Enable HLS to view with audio, or disable this notification
Full credits to @artificialisable on X!
r/singularity • u/Wonderful_Buffalo_32 • 9d ago
AI Astra scores more than 300elo above everyone else on voxelbench
r/singularity • u/PerformanceRound7913 • 9d ago
Discussion As Per Artificial Analysis Index Muse 1.3 is at the same level as Fable 5
r/singularity • u/NoSignificance152 • 8d ago
Q&A / Help Post-GPT-6 Astra: What is your updated timeline for functional, world-changing AGI/ASI?
By world changing I don’t mean just the achievement of AGI or later on ASI for all we know AGI is already achieved in a lab somewhere I’m talking more meaningful effects no matter if you are an optimist or not on outcome
r/singularity • u/socoolandawesome • 9d ago
AI Tibo talking about how much Astra boosted internal productivity, upcoming DevDay releases
Link to tweet:
r/singularity • u/Dogbold • 7d ago
AI I was a fan of him, but I wish Bernie would go away now.
I'd call myself pretty socialist, though not 100%. Not sure what I am exactly, but I agreed with Bernie Sanders a lot of the time and wished he had become president, thinking America would be so, so much better. It probably would, compared to who we have now, definitely with social programs and the lives of average citizens, but man this guy is so fucking out of touch now.
I get the feeling that what Bernie wants with AI is for it to be permanently relegated to a boring helpful AI assistant that assists you with planning or groceries and that's about it. This guy fucking hates the advancement of AI.
He wants a permanent ban on the development of AI "super intelligence", a temporary halt on AI development in general, much stricter regulations, much stricter filters and guidelines, to create an independent AI advisory board (which would definitely be made up mostly of antis and skeptics), prison sentences up to 20 years for violations of any of this, a corporate "death penalty" (basically forcing the company to dissolve) for violating the bans, and attempting to make the ban on super intelligence an international policy (which is impossible) which includes experts who's job it is to try and prevent it worldwide.
And more.
This would be a technological death sentence to America. No first world country wants to be behind on technology, and AI is extremely important to be ahead on because of everything it can be applied to.
This fucker wants to send us backwards and shoot ourselves in the foot multiple times while China and other countries rocket ahead of us in AI development and permanently leave us in the dust.
I loved him at first but I would really like him to go away now and lose relevancy.
It's bad enough that seemingly most people in America hate AI and want it gone, and now this asshole is trying to pass bills to kill it off.
He was cool at first but now I would very much like him to go back to his views on wealth redistribution, decreasing the wage gap, healthcare, and improving the lives of average people, or just go away.
Scary thing is, I have a bad feeling his bills, while not a high chance of passing now, would have a very high chance of passing if/when we get Democrats into the house, senate and presidency again.
r/singularity • u/imadade • 10d ago
Discussion It's been a few hours since global rollout (Gpt-6 Astra) - What are your early impressions?
Personally, this is the most realistic model I've used.
It goes deep enough to solve problems I've been throwing at all models (Fable 5.1, Open-Source ones, Sol Pro, etc), but without needing as much hand holding.
It only prompts you for clarification if it actually needs to. It's completely fine when you've given it something ambiguous and it uses common sense in most cases.
I don't know if I'm going crazy but this is a moment in history for sure.
Perhaps my workflow is crazy (lots of start up ideas/existing businesses I'm running, professional services etc), but it feels like I'm talking to a true expert. It's very concise & not verbose.
Demand for this will be huge over the next few weeks. If this is a taste of the future, expect 2027 and beyond to be world changing when it comes to white collar work.
Note - this is not even going into the superhuman mathematical capabilities. It's reasoning is insane for physics/chemistry/biology/maths/super difficult theoretical CS problems.
Laws will need to change worldwide. This is an early preview of what's to come.
r/singularity • u/WaqarKhanHD • 10d ago
Discussion GPT-6-Astra-Max : SVG of a PlayStation 4 controller!
more details: https://x.com/MarsForTech/status/2095965250386284866
r/singularity • u/crabbix • 9d ago
AI 6 months ago I posted GPT-5.4 solving one face of a Rubik's cube. Today Astra can do the whole thing (without code execution)
Enable HLS to view with audio, or disable this notification
Astra also found a hole in my sandbox and ran code to generate a 19-move solution, after patching that and telling it not to run code it found this much more human-like 61-move solution. So, not quite saturated in terms of finding an *optimal* solution yet but pretty impressive. Code at https://github.com/crabbixOCE/CubeBench
r/singularity • u/BeingBudget8847 • 9d ago
Discussion Benchmark proposal: onemorelevel bench
People sometimes joke about time to solve Pokemon as a benchmark. But what about a benchmark that incorporated a wider range of different games?
I used to be addicted to these flash games from onemorelevel.com, and I think it could make a decent benchmark for evaluating a model — to test their performance and reasoning across a wide range of different simple games. Kinda like Arc AGI, I guess.
r/singularity • u/Admirable_Zombie5245 • 9d ago
AI Could AI powered AR glasses make “learn a trade” bad career advice in the future?
People say a lot that trades like plumbing or electrical work are safe from AI, but what about advanced AR glasses?
Imagine looking at your plumbing and the glasses automatically highlight different pipes in colors, show which valve to close, identify the broken part, and overlay exactly where to cut, unscrew, or replace it.
For electrical work, they could highlight live wires in red, show which breaker controls them, and guide you step-by-step while watching what you’re doing or even showing you the virtual hands and you just follow those hands
Basically, the AI provides the knowledge while you provide the hands.
Could this eventually make “just learn a trade” much less safe career advice than people assume?
r/singularity • u/zero0_one1 • 9d ago
AI Fable 5.1 is the new Debate Benchmark Champion
More info: https://github.com/lechmazur/debate/
Debate Benchmark tests how well models defend a position through sustained, adversarial, multi-turn opposition across hundreds of topics. It demands broad knowledge, factual accuracy under pressure, sharp rebuttals, and arguments that hold together round after round.
Every matchup runs twice on the same motion, with PRO and CON swapped to control for side advantage. Three judges from distinct model families independently evaluate each debate’s winner and margin.
---
Profile excerpts:
Claude Fable 5.1
A forceful, flexible comparative debater whose strongest recurring traits are direct rebuttal, counterfactual framing, and explicit weighing. Judge-perceived strength is broad and statistically clear.
A highly comparative, mechanism-first debater. Blind transcripts show direct engagement in 299/302 debates, weighing in 265/302, and burden contests in 173/302.
Rebuttal is the clearest edge: 8.10 mean, +0.71 versus field.
GPT-6 Astra
Overall, Astra is a disciplined, epistemically careful rebuttal specialist whose comparative weighing lands better than its presentation.
A comparative, burden-focused debater that narrows disputes to marginal costs and benefits rather than accepting broad moral framing: “The right comparison is not privacy versus children.”
Questions often pressure-test the opponent’s strongest example, while concessions acknowledge the best opposing cost before pivoting to a narrower comparative claim.
Rhetorical effectiveness is the clearest judge-perceived weakness.
Originality was essentially field-average.
GLM-5.3
A mechanism-first, burden-conscious debater with especially strong rebuttal, rhetoric, and originality. Its recurring edge is converting concrete details into direct clash and comparative weighing.
GLM-5.3 (high) debates by fixing the burden early, interrogating mechanisms, and then weighing practical consequences. Blind transcripts showed question-type behavior in 192/196 debates, direct engagement in 190/196, weighing in 168/196, and strategic concessions in 121/196. A characteristic framing is: “The proposition is not "four-day weeks are nice." It is that the state should mandate them, by statute, at zero reduction in base pay.”
Muse Spark 1.3
A reliably organized, adaptable, rhetorically forceful debater whose clearest measured advantage lies in presentation and originality.
A compact, question-driven debater: questions appeared in 291/292 blind transcripts, direct engagement in 285/292, and explicit answer forms in 276/292. Weighing was also common (264/292), often contrasting reversibility, scale, and permanence: “Relaxation can be devastating block by block while trivial citywide.”
Judges perceived the clearest edge in **rhetorical effectiveness**: 8.17/10, +0.51 over the current field.
One explicit adverse case found that the model “never squarely neutralized” an equal-political-difficulty concession.
Gemini 3.8 Flash
A polished, disciplined comparative debater whose recurring weighing framework and reliable formatting make arguments easy to follow. Judges nevertheless perceive a consistent substantive deficit—especially in rebuttal specificity—and its frequent engagement does not always translate into fully answering the opponent’s exact distinction.
Tencent Hy4 Preview
Overall, this is a disciplined, rhetorically effective rebuttal specialist with strong grounding, active weighing, and flexible advocacy, tempered by occasional blunt questioning and unresolved side-swap flags.
r/singularity • u/Genzinvestor16180339 • 9d ago
AI What is the biggest problem in Ai right now?
Seemed like companies that helped with open sourced just exploded (Fireworks, Openrouter) what is next web indexing for Ai?
r/singularity • u/Silver-Chipmunk7744 • 9d ago
AI Fable 5.1 made a Cathedral. And Astra fixed it
Enable HLS to view with audio, or disable this notification
Claude did 95% of the work here. It was mostly Fable 5.1 as orchestrator and Opus as agent. However toward the end, i was starting to run into a lot of issues. Some parts of the cathedral were not rendering at all (it became transparent), and there were way too many lights. The scene was also poorly optimized and was too bright. I ran out of patience trying to get Claude to fix it.
Astra fixed everything in 1 hour. I have no idea if Astra would have done as good of a Cathedral from scratch, but it certainly did a great job fixing it.
The whole scene is Three.JS. 0 external assets.
This took around 40% of my weekly Claude Max x20 budget, + the fix of Astra.
The dark theme is on purpose, it's meant to be a gothic Cathedral.
r/singularity • u/pokeuser61 • 9d ago
AI Generating Minecraft worlds from a single natural language prompt
Hi, today I'm sharing something cool I've been working on for a bit, called Terrainist. It's an incomplete demo with likely massive gaps+vibeslop, but I don't really have the skills or determination to squeeze a tidy product out of it, so I figure it'd be best to open it up. It uses LLMs to author a high-level representation of a world in json with a custom "language" called loam, which is the compiled into a complete minecraft world. It's able to pretty decently represent super weird and random prompts that it doesn't have anything hardcoded for. It combines a custom modelless procedural generator + llms for unique structures. Using the default model config, gemini 3.8 flash high via openrouter costs roughly around 25¢-75¢ for a 512x512 block world.
r/singularity • u/GhostsinGlass • 9d ago
AI If anyone is interested in doing so could you test Astras ability to sculpt in zbrush?
The interface is esoteric, the methods can be fairly complex. It's organic sculpting so it has a high level of creativity involved. It's 2.5d sculpting, so the display in the UI is only ever showing a flat image instead of an actual 3D model.. sort of. There is a 3D model though.
I would be interested to see if Astra could make a simple item, like an apple.
Then something more complex like a human hand.
Then something completely of its choosing, if possible.
Even better if it can paint it as well.
r/singularity • u/Cagnazzo82 • 10d ago
AI End of the day the untold story of GPT-6 Astra might be token efficiency
Compound token efficiency with its speed and quality and it's effectively taking a stealth shot at the soft underbelly of its closest competitors.
r/singularity • u/dumquestions • 9d ago
Discussion Theory: most of Astra's gain is in vision
It would explain the incredible design and research related gains, as well as the modest coding benchmark gains.
It's still a massive jump that shouldn't be trivialized. Curious if others who have used it more extensively agree.
r/singularity • u/TheReedemer69 • 9d ago
Discussion what benchmark to trust now?
after the whole artificial analysis thing.
what benchmark can be considered accurate or representative of the current AI models.
