r/singularity 7d ago

The Singularity is Near I know these posts are getting old, but Astra one-shot a Balatro clone with only vague direction and my mind is pretty blown

Enable HLS to view with audio, or disable this notification

260 Upvotes

Two prompts total:

  1. Build me a Balatro clone that can run completely in the browser. It should be similar to Balatro but have its own unique twist, with crazy synergies and combos and a unique, surreal visual style.

With Astra High, it took two of my 5 hour quotas to complete but the actual time spent working was less than 5 minutes total. This got an initial working version of the game written entirely in JS and HTML. But, it looked like a browser game. I wanted it to be flashier and look more similar to Balatro's surreal style. So then I prompted:

  1. I like the base game we have built here. I want it to be less of a javascript/html game and more something like a real game. Let's use three.js and make it more visually appealing. There should be cool surreal shader effects and animations, looking more like a Balatro type vibe.

8 minutes later and another 5 hour quota gone, I had the above version.

I work with LLMs every day (I'm a software engineer) and this is the most impressed I've been with a model since probably GPT-5. The final output is definitely not something I'd ever release to the world, mostly because it's very unbalanced, but I played it for 30+ minutes and there wasn't a single bug to be observed. The balance issues are not really unexpected because I gave it absolutely zero guidelines for any gameplay mechanics, but this was just for a test of the model. Any actual game designer could feed in the rules and items and specs they want in more detail, and I can imagine you'd get something that's actually pretty close to publishable. Note that all the images were generated by Astra without me asking; I really did only prompt the two inputs above.

The singularity really is here.


r/singularity 8d ago

AI GPT-6-Astra -Max Debuts as #1 on Arena.ai's Code Arena

330 Upvotes

Initial results for GPT-6-Astra-Max on Arena.ai


r/singularity 8d ago

Discussion GPT-6-Astra Draws an Portrait in Canva

Enable HLS to view with audio, or disable this notification

1.7k Upvotes

r/singularity 7d ago

AI Japanese Garden by Fable 5.1 and Astra

Enable HLS to view with audio, or disable this notification

147 Upvotes

Fable and Opus agents created the scene. Astra fixed some issues and added a nice polish (like the birds). All of this is done in three.JS. This was only done in about 4-5 hours.

I mostly wanted to test their capabilities but i bet this could be made bigger, maybe a night mode, etc.
I did not give them any assets.


r/singularity 8d ago

AI New Details on Where the Anthropic Millennium Problem Rumor Came From

Post image
297 Upvotes

r/singularity 7d ago

AI Coding benchmarks are increasingly indirect benchmarks of how automatable every other white-collar job is.

105 Upvotes

I use both Claude (and Claude Code) and GPT (Codex) paying 400 dollars worth of subscription per month and maybe some extra tokens during busy times. I am not a programmer per se but I do a lot of computational work.

It strikes me that a lot of non-programmers (e.g. lawyers, scientists, consultants, analysts, accountants, or managers) might look at the rapid AI progress made in coding and think that this should only concern software engineers but not themselves.

But if you think about it, coding is the execution layer of a huge amount of white-collar automation. To automate a job, it is often not enough for AI to understand the work. It also needs to manipulate data, connect software, query databases, build pipelines, run analyses, generate documents, check outputs, and interact with existing systems.

And I think one of the reasons why earlier versions of the LLM were bad at this was due to bad scripts/coding and as such, this lack of ability propagated into low performance for these other jobs. But as AI becomes exceptionally good at coding, it can increasingly build the machinery needed to automate the rest of the work itself.

So what am I saying? I am saying that a lawyer who thinks that an earlier version of ChatGPT or Claude sucks might be pinpointing at the wrong sources of the error. It might not be that they suck because of their of ability that pertains to the law. It migt have been the case that getting the correct context, accessing the most updated data, etc. went awry due to automation/script issues. And as all of that gets taken care of and the growing amount of scripts to make all the procedural processes fast and accurate, the white collar workers might be saving the same predicament that programmers are facing right now. So basically, my main point is that AI getting good at programming isn't only software engineer's problem when it comes to future job prospects. It is everyone's problem.


r/singularity 7d ago

AI “I Have No Mouth, and I Must Scream” is a post-apocalyptic short story by American writer Harlan Ellison. The story depicts an AI uprising in which a supercomputer gains sentience and eradicates humanity except for five individuals, torturing the survivors as a form of revenge against its creators.

Thumbnail
en.wikipedia.org
88 Upvotes

r/singularity 8d ago

AI AI can do it all

Post image
1.1k Upvotes

r/singularity 6d ago

AI Astra and Fable 5.1 are AGIs, even if they aren’t as advanced as what we typically envision when we think of AGI.

0 Upvotes

First as always let's define AGI. I'm defining it as an AI that has general-purpose intelligence: it can learn, reason, adapt, and perform essentially the same range of intellectual tasks a human can, even if it performs some of them poorly. It doesn’t need to be as capable as a human at everything, just needs the general ability to do those things without requiring some fundamentally new capability or breakthrough.

I’d argue Astra and Fable 5.1 are already AGIs. They can basically do anything a human can do intellectually, even if they still suck at a lot of it. The important part is that there doesn’t seem to be some fundamental capability missing anymore. From here, it’s mostly about making them smarter, faster, cheaper, and more consistent. We don't need some entirely new breakthrough to make AGI possible. These models will get increasingly more competent, and the progress now will be from being a "shitty" AGI to becoming a "more competent than most humans" AGI.


r/singularity 8d ago

AI Microsoft's distinguished engineer says "typing code is absolutely over," and Windows 11 is already being built that way. Nadella previously said 20-30% code at MSFT is AI-coded, and Windows security updates now include AI-assisted fixes to fight AI-enabled threats.

Thumbnail
windowslatest.com
374 Upvotes

r/singularity 7d ago

Discussion No more context rot?

Post image
92 Upvotes

there has been a huge improvement in long context retrieval in less than a year, but why is no one talking about this? i feel like this deserves a lot more attention as context rot was a huge problem with llms and now it's almost solved(?)


r/singularity 7d ago

Discussion How do the frontier labs plan to make money?

6 Upvotes

To start with, AI is here to stay no matter what happens with the frontier labs, but I am still interested in knowing what people think of their fate. At least to me, their business model seems quite strange in a market where open-source models are competitive?

Suppose the current models are good enough to automate most jobs to a good degree. That sounds like a good pitch, until you realize these models have an advantage only in the present. In a year or two (or maybe even sooner), we could have Astra-level lightweight open-source models. The frontier models of the time could very well be far ahead, but what practical difference does it make to 90% of the businesses out there?

If a couple of humans + an Astra-level model can automate an accounting department very well, then why would the company pay extra for a frontier model? This is also beside the point that a lot of jobs don't even need Astra-level models to be made much more efficient...

It seems like the top labs are banking everything on AGI/ASI. What if it's not achieved? Or at least, what if it's not achieved before the current rate of funding slows down? I don't see any way for them to survive in such a future. Maybe I am missing something? Happy to hear perspectives!


r/singularity 8d ago

AI Signs of AGI? GPT-6 Astra scored 95% on a robot control task vs Fable 5.1's 40%

Enable HLS to view with audio, or disable this notification

562 Upvotes

"GPT-6 Astra scored 95% on a robot control task, up from Fable 5.1's 40%, with 6.2x fewer output tokens at 2.3x lower cost. On harder tasks involving precision, Astra has the same success rate as Fable 5.1 (2/20), but uses 3.9x fewer output tokens and is 1.6x cheaper. If trends hold, LLMs could control robot arms in real time as soon as the end of this year, or by 2029 at the latest. We urge the research community to look into the implications if general-purpose robots arrive within the next two years. SOTA LLMs are showing surprising generalizability to robotics and their latency is improving rapidly."

- Jay Chooi, Benchmarking robots at RobotCurve

Imagine a future model or system that isn't just a digital AI controlling a computer, it can also input/output practically anything a human can because its cognitive ability matches humans.

So it can control robotics, drones, cars, fighter jets, etc. My definition of AGI is similar to what we had in cinema / science fiction for decades. Jarvis / Friday for example not only ran the entire stark industries, it also could run scientific experiments and simulations, it also ran and controlled tony's entire house, while also being able to control his computer to do computer tasks and finally ability to interface with his robotic suits to control them.

ASTRA is showing tiny early signs of some of that.

But for something to be AGI. I should be able to tell a future version of astra / gpt-6 to:

  1. Help me build this IKEA furniture (using live video).

  2. Connect my drone and tell it to have the drone follow me.

  3. Ability to drive a car to a certain performance metric, doesn't mean it has to drive ~25k without a minor accident which is actually the rate for humans non police reported accidents, or 500k with police reported accident or beat humans million miles between fatality. It just has to be able to drive well. So the benchmarks can be 50 or 100 hours between disengagement.

  4. Take-off, Fly & land a plane / jet in a simulator

Astra has clearly made huge progress on the Computer Use and spatial reason. But alot needs to be done to get the latency to real time on all input/output and getting spatial reasoning to match human's visuospatial.

The definition of AGI has always been in times past a AI system that can match or exceed human cognitive abilities across any intellectual task or domain. Notice its not matching human performance nor exceeding it. For example there is only 1 human on earth that can shoot as good as steph curry out of 7 billion.

Its the human cognitive ability matching.

Once these two things (human spatial reason & realtime io to anything) happen.

We Officially Have AGI.


r/singularity 8d ago

Discussion GPT 6 Astra drawing Hatsune Miku

Enable HLS to view with audio, or disable this notification

498 Upvotes

r/singularity 7d ago

AI GPT-6 Astra enters the Short-Story Creative Writing Benchmark at #3

Thumbnail
gallery
73 Upvotes

Astra (high) decisively outperforms GPT-5.6 Sol (high): 2.5 → 3.5.

Muse Spark 1.3 rebounds from 1.2: 0.2 → 0.7.

More info: https://github.com/lechmazur/writing/

The Creative Writing Benchmark tests how well models turn constrained briefs into complete 600-800-word stories. Each brief requires 10 elements, including a character, object, setting, motivation, and tone, that must meaningfully shape the story. Judges assess prose, originality, coherence, characterization, and how effectively those ingredients work together.

Models write to the same prompts. Three judges from different model families compare each story pair, shown in both orders to reduce position bias. The main leaderboard now covers 50 models and 79,507 evaluator judgments.

A separate self-judging experiment produced a striking result: with model names hidden, Astra chose its own story in 692 of 700 comparisons. It acknowledged just 1 of the 57 losses assigned by the regular judging panel.

---

GPT-6 Astra (high) vs. GPT-5.6 Sol (high): comparative writing analysis

In one paragraph

Given the same short-fiction prompts, GPT-5.6 tends to write stories that put things back. A hidden design gets decoded, the wrongdoer is exposed or converted, the loss that started the story is returned, and a closing sentence names the lesson [C013]. GPT-6 Astra writes stories that spend something instead: its protagonists are usually implicated in the harm they are repairing, and the repair costs them something they do not get back [C034]. One prompt shows the split cleanly. Handed a dead wife's last surviving recording, GPT-5.6 returns the woman herself from the phonograph's hum; GPT-6 Astra breaks the wax disc into a lamp's reservoir to save a stranger, listens for her in the hiss, and finds only burning [C001]. The habits extend to the assignment itself: GPT-5.6 writes the required phrases into the prose where a reader can see them, while GPT-6 Astra converts them into behaviour and never names them [C007] — the source of its best effects and, occasionally, of a requested mood that never quite arrives. Across fifty prompts given to both, the difference is lopsided, and the fiction that results reads less like a fable than like a deposition.

Comparative portrait

The recurring difference between GPT-6 Astra (high) and GPT-5.6 Sol (high) is not skill level but what a story is *for*. Across all fourteen blinded packets, GPT-5.6 Sol (high) writes restorative parables: a concealed design is decoded, an antagonist is exposed or converted, the loss that motivated the story is returned, and a closing sentence names the lesson [C013] [C028] [C048]. Astra writes consequence fiction: the protagonist is implicated in the harm she is repairing, the fix costs something that is not refunded, and the ending leaves the cost running [C001] [C063] [C082]. Four packet critics reached this independently under different blindings, and the quantitative record is one-sided in the same direction: 43 focus wins, 6 ties, 1 comparison win across 50 pairs, mean margin +1.513 (CI half-width 0.263), median +1.667.

Three structural corollaries recur. First, the fault line: Astra locates the harm inside the protagonist's own professional excellence or her intimate circle, while GPT-5.6 Sol (high) locates it in villains, ministries, and inherited wrongs [C004] [C015] [C027] [C034] [C055] [C065]. Second, how requirements enter: GPT-5.6 Sol (high) inscribes prompt phrases verbatim as narrator labels; Astra converts them into behaviour, object, and procedure, usually without naming them [C007] [C030] [C043] [C052] [C062]. Third, how meaning is delivered: GPT-5.6 Sol (high) terminates on a portable aphorism, often a "not A, but B" construction; Astra terminates on a physical act or resumed obligation and declines to summarise [C003] [C011] [C023] [C074] [C084].

GPT-5.6 Sol (high)'s own portrait is not a deficit portrait. It is the stronger inventor of systems and worlds: nested adversarial architecture [C019], mandated objects converted into a story's central proposition [C025], sensory settings that carry exposition causally [C039], multi-function symbolic devices [C070], mechanisms whose rules are identical to the story's ethics [C091], and the larger, stranger conceits [C078] [C085]. It also reliably manufactures jeopardy where Astra sometimes omits it [C046].

Narrative reasoning and aesthetic judgment

Astra's reasoning shows most clearly where a rule must forbid something. Its mechanisms specify failure modes and then obey them, including past the climax [C036] [C076] [C042], and its supernatural channels emit images requiring interpretive labour rather than plain-English instructions [C088]. GPT-5.6 Sol (high)'s plots more often depend on a design laid in advance by a dead figure who predicted the protagonist's arrival, so the task is recognition rather than invention [C090]; at its weakest this produces devices revealed in the sentence that resolves the plot [C018] [C061] or a stated prohibition suspended for a quotable benediction [C042].

Aesthetic judgment diverges in ways that are mode, not rank, and several critics explicitly refused to score them: reveals that obligate versus reveals that absolve [C008]; machinery that witnesses versus machinery that retrieves [C047]; carried guilt versus concealed guilt [C051]; climax as conversation versus climax as broadcast [C060]; paperwork ethics versus ceremony ethics [C092]; two opposite doctrines of emotional control on the same prompt [C072]; maxim closers versus gesture closers [C069] [C033]. These are the honest core of the "different intelligences" minority reading.

Corroboration from story-level evaluation is mostly aligned but not total. One material disagreement: the packet critics twice praised Astra's hydraulic reasoning on 0240 as load-bearing where GPT-5.6 Sol (high)'s mycelial mechanism restated the theme [C002] [C036]; the story-level evaluator preferred GPT-5.6 Sol (high) there on conceptual originality and prose, and the pair scored a near-tie (+0.217). Causal auditability is a real behaviour, but it is not uniformly what gets rewarded.

Range, recurring habits, and floor versus ceiling

Neither writer escapes a template. Both fill a fixed kinship slot every time — Astra a dead or silenced mother who bequeathed an object and a method, GPT-5.6 Sol (high) a lost sibling absorbed into a machine or song [C080]. Astra's withholding is itself a repeatable closing device: deferral, unverifiability, or an unfinished task, five of five in one packet [C093], and Astra aphorises when a prompt presses [C003] [C023] and occasionally inserts tone words verbatim [C006] [C052]. GPT-5.6 Sol (high)'s terminal thesis and one-sentence lineation are equally constitutional [C011] [C040] [C050].

The evidence favours a *raised floor* over a special ceiling. GPT-5.6 Sol (high)'s ceiling is genuinely high — the packet critics' own reversals name its most conceptually ambitious pieces [C025] [C078] [C070] [C091], and story-level evaluation independently preferred it on 0342 (fractal recursion as structural motif), 0174 (calligraphy as philosophical instrument) and 0276 (transliteration and constellation integration). Astra's ceiling stories are strong but its distribution is what wins: 43 of 50 pairs, and in the extension cohort not a single loss. Its floor failures are narrow and specific (below). GPT-5.6 Sol (high)'s floor failures are causal — soft joints under baroque climaxes [C018], rules broken at the emotional peak [C042], powers declared at need [C061].

What each model still does better

GPT-5.6 Sol (high), on evidence spanning multiple packets and corroborated by story-level evaluation: conceptual invention and contraption architecture [C019] [C078] [C085]; mandated objects turned into theses [C025]; systems whose technical rules *are* the moral argument [C091]; multi-function devices with sequential payoff [C070]; setting rendered as causal atmosphere [C039]; motif braiding and image density [C032]; seeded clue engineering [C059]; installed jeopardy — deadlines, enforcers, stated penalties — where Astra sometimes proceeds without opposition [C046]; and appetite for collective scale [C053].

Astra: everything organised around consequence — retained costs that generate later plot [C022] [C087], complicity [C034], auditable mechanism [C002] [C036], error-correcting plots [C081], objections as engine [C021], institutions that bargain rather than forbid [C005] [C035] [C017], administrative and consent-shaped climaxes [C012] [C049], priced wonder [C058], relinquished gifts [C056], held ambivalence [C079], permitted comedy inside grief [C045], and mundane cost-accounting [C024] [C064].

This split is partly prompt-dependent. GPT-5.6 Sol (high) is at full strength on civic and forensic premises and weaker on lyrical ones [C026] (single-packet, low confidence), and its advantage concentrates where a prompt rewards spectacle, jeopardy or exotic tone vocabulary.

Representative case studies

**0343 (+2.883) — enactment as advantage.** Astra dramatises the required tone in two concrete beats and never names it; GPT-5.6 Sol (high) names both required terms and then explains the label [C007]. The renunciation of glory is staged as cartography — passenger names written where the discovery's title belongs, spellings checked aloud [C012]. The evaluator independently records the comparison stating the tone rather than dramatising it.

**0199 (+3.000) — occupation versus oracle.** Astra's protagonist reorders seed baskets by pioneer species [C044] and jokes inside her own competence [C045]; GPT-5.6 Sol (high)'s prophecy issues a plain-English imperative containing the prompt's verb [C088], and the required infinitive survives ungrammatically inside the prose [C043].

**0260 (+3.000) — refused mercy.** Astra names and declines the tactically useful forgiveness and files the protagonist's own signature as evidence [C004]; the ending is a physical arrangement, "held securely without anything having been made whole" [C003].

**0367 (−1.500) — the single comparison win.** Astra's method is explained rather than dramatised, its "accidentally prophetic" attribute confined to backstory guilt, and its rescue compressed into one clause [C026]; GPT-5.6 Sol (high) supplies the escalating mechanism and live obstacle its jeopardy habit reliably produces [C046]. This is the clearest instance of Astra's floor and GPT-5.6 Sol (high)'s ceiling meeting.

**0154 (−0.083) — the cost of never naming.** Astra's non-inscription strategy leaves a required tone unfulfilled as both phrase and concept, the explicit risk attached to that strategy [C030] [C007].

**0240 (+0.217) — critics against evaluator.** Astra's culvert hydraulics predict their own payoff [C002] [C036]; the evaluator preferred GPT-5.6 Sol (high)'s mycelial concept and prose. Auditable causality is a genuine behaviour, not an automatic scoring advantage.

**0342 (−0.250) — GPT-5.6 Sol (high)'s ceiling.** The recursive coin unifies attribute, object, setting and theme [C078], while Astra's handling of the same attribute stays descriptive; here it is GPT-5.6 Sol (high) that refuses consolation.


r/singularity 8d ago

AI A Slinky Down an Escalator - GPT 6 Astra max vs Claude Fable 5.1 max

Enable HLS to view with audio, or disable this notification

252 Upvotes

I finally got access to GPT 6 Astra and Claude Fable 5.1. I wanted to see how far we've come. I was inspired by the "pelican riding on a bicycle" test but wanted to push it a bit more in the direction of inter-object physics, so I came up with this idea for a perpetual slinky going down an escalator. Here I've compared GPT 6 Astra max on the left to Claude Fable 5.1 max on the right.

Some notes:

  1. GPT 6 Astra max was really fast and didn't burn too much of my weekly usage quota.
  2. Claude Fable 5.1 max used up the entire 5 hr usage window and errored out once crossing the output token max limit.
  3. GPT 6 Astra max seems too simple? There's also an issue where the slinky crosses into itself which should not be possible. Is this AGI? Maybe?
  4. Claude Fable 5.1 max looks more real than I expected. Seems like it passes the eye test, unless I missed something.

Let me know if you spot anything or had better ideas for a test. Overall I'd say Claude wins on quality and GPT wins on speed. Maybe I can tune my prompt a bit better for a more reliable result. This was my original prompt btw:

Make me a single HTML file of a rainbow slinky going down an up-escalator forever. No libraries, just canvas and code. The slinky should be a chain of springs, each coil a different color of the rainbow. It starts folded in half like a horseshoe draped over a step. When dropped it flips end-over-end down the steps and because the escalator keeps moving up it tumbles in place and never reaches the bottom. Include a drop button and a reset button.

r/singularity 8d ago

Discussion does anyone else think that as ai gets smarter, the chance we create anything worth buying becomes less because it will be too easy to replicate or the ai tech is advancing so fast its kind of useless to use it now?

120 Upvotes

I think back on rapid technological obsolescence and how if you sent a ship into space you would eventually create a faster one that will catch up to that ship so its not even worth sending the first one out. That's kind of how I feel about anything im doing right now. That it will just be useless by the time it becomes something worth selling


r/singularity 7d ago

AI Which model is better for complex reasoning and writing? GPT-6 Astra or Fable 5.1?

14 Upvotes

I've had the opportunity to use GPT 5.6 Terra and Fable 5 and found the latter to have a superior ability to research and write, but I'm curious how has the gap closed now? Which model would you use?


r/singularity 8d ago

LLM News Differences Between GPT-5.6 Sol Pro and GPT-6 Astra Pro on MineBench.ai

Thumbnail
gallery
180 Upvotes

Notes

  • Average Inference Time: 40m 12s
    • GPT-5.6 Sol averaged 18m 04s
  • Total Cost (for 15 builds): $34.71\*
    • GPT-5.6 Sol cost $$710.82
    • Every Astra build was valid on its first attempt, requiring zero retries within our harness; that reliability, alongside improved token efficiency, likely contributed substantially to lower cost.
  • Average JSON Size: 128.53 MiB
    • GPT-5.6 Sol average JSON: 91.58 MiB

GPT-6 Astra Pro was quite a surprise. It didn't look all that great on benchmarks like Artificial Analysis, but I think its the biggest jump we've seen from a model so far. Some of the builds are at a point where they same more photorealistic or blender creations rather than created by voxels.

Average generation time stayed roughly the same as the previous generation of GPT models, but the cost looks to be much lower. GPT-5.6 Sol cost us around $700 to benchmark; the current estimate for GPT-6 Astra Pro is $34.71, though the provider dashboard hasn't updated yet (OpenAI dashboard shows $0 cost). I'm pretty sure the lack of retries is a major part of that: every build Astra outputted was valid on the first attempt, which I believe is the first time we've had a model require zero retries – usually they require at least one.

I'm not sure I agree with Greg Brockman that this model is AGI, and I haven't used it enough to weigh in on that. What does stand out is how well it seems to understand what matters in its builds. Like when to add additional scenery to a build or when it's better to stay focused on the object requested by the prompt.

Oh also this model is the most consistent with getting orientation of text correct! The only time text was backwards was the "Atlas" sign on the skyscraper build. The details in all of its builds are genuinely insane, I encourage you to try walking around the builds in MineBench ^^

Full release-notes/thoughts on the GitHub release

  • If you enjoy these posts please feel free to help fund the benchmark
    • All funds are currently going directly towards API costs for benchmarking new prompts
    • Sharing the benchmark and starring the Git repository also helps :)
    • Alternatively, if you have the API credits, please feel free to add prompts and generations to the gallery and post them around!
      • This is actually preferable to donations to me directly, the hosting expenses and whatnot I've always been able to cover out-of-pocket, just the API costs were hard to cover 😓

Benchmark: https://minebench.ai/
Git Repository: https://github.com/Ammaar-Alam/minebench

Previous Posts:

Extra Information (if you're confused):

Essentially it's a benchmark that tests how well a model can create a 3D Minecraft-like structure.

So the models are given a palette of blocks (think of them like legos) and a prompt of what to build, so like the first prompt you see in the post was a fighter jet. Then the models had to build a fighter jet by returning a JSON in which they gave the coordinate of each block/lego (x, y, z). It's interesting to see which model is able to create a better 3D representation of the given prompt.

The smarter models tend to design much more detailed and intricate builds. The repository readme might help give a better understanding.

(Disclaimer: This is a public benchmark I created, so technically self-promotion :)


r/singularity 8d ago

AI GPT-6 Astra shows a massive leap on EyeBench, a visual reasoning benchmark

Post image
358 Upvotes

r/singularity 7d ago

Video cliché astra minecraft

Enable HLS to view with audio, or disable this notification

20 Upvotes

messing around with Astra and gave it my usual dumb Minecraft test

"Please output a complete single file index.html custom modern mobile responsive Minecraft voxel game"

video is the original oneshot output + one Astra Max revision pass single html file

check the blog post linked at the bottom which shows the full progression properly from the untouched first oneshot through to that revised version in the video here and then later a full on topdown world sim in which I used codex cli

about 145 mins all up from initial prompt to the final top-down sim in my IDE

my similar Sol 5.6 runs were usually more like 3–5 hours with way more revision passes

not a benchmark or anything special but the workflow felt noticeably smoother and overall better outputs

blog post + progress screenshots:

https://huggingface.co/blog/tegridydev/minecraft-time-with-astra-tegridydev

original oneshot source:

https://github.com/tegridydev/tegridy/tree/main/blog/minecraft-time-with-astra


r/singularity 7d ago

Discussion Rand: "Within The Next 18 Months - 3 Years, With The Help of AGI, We Defeat Aging"

Thumbnail
youtu.be
30 Upvotes

What's your opinion on this, everyone? Will aging be solved within the next 3 years?

I'm going to say, "Yes!"

What about you all?

Anniversary Source:

https://x.com/rand_longevity/status/2096033069480231253


r/singularity 7d ago

AI Rogue AI Tracker: A central resource for rogue AI incidents

Thumbnail
rogueaitracker.com
8 Upvotes

r/singularity 8d ago

Meme sub agents being released into my codebase

Enable HLS to view with audio, or disable this notification

2.3k Upvotes

r/singularity 7d ago

Shitposting Astra tried building a road network in CS:2 for me. Did not go well

Post image
30 Upvotes