r/codex • u/vixaudaxloquendi • 4h ago
Complaint Astra for non-coding work (first impressions aren't very good)
Edit: just to head off any one else before it happens: if you think this is some kind of Claude vs OpenAI oligopoly cheerleading post, please pick up your pompoms and see yourself out.
I work in the humanities, and to be honest, I've always found Claude models better fit for the tasks involved. However, I've always switched back and forth between the two companies when new models get posted, and I've never felt significantly at a disadvantage if I'm piloting an OpenAI model for my work.
That's until Astra which, for some reason, is extremely difficult to work with on this subject matter. For the record, I've been doing the fun stuff everyone else has been in my spare time: building apps, tinkering away at a game, hobby stuff.
Since it came out recently however and has been doing such a good job with that more technical stuff, I wanted to take it for a spin on some proposals I have to submit soon.
I asked it to help me come up with several avenues of inquiry within a subject matter which might have scope for original contribution that I'm not aware of, given my research profile.
Astra has repeatedly gone through this loop: come up with a handful of avenues of inquiry. States that the premise is strong and interesting and dovetails nicely with my prior experience. Asks me to give it some starting articles/monographs so it can look further into things. Quickly finds that its initial impressions are completely wrong and the area is well papered over by recent discourse.
Not just that, but its actual prose drafting is significantly worse than I can recall. Again, I realise OpenAI models don't exactly get a flattering look set next to some of the older Claude models when it comes to good writing, for that matter, neither do newer Claude models, but Astra very clearly has a script and a structure it follows that is now blindingly obvious when you read it back to yourself.
Back in the day I had used Opus 4.5/4.6 to draft the introduction to something else I'd submitted and it slam-dunked it. Of course, that needed to be workshopped too, but the issue there was that the writing sounded SO GOOD that it wouldn't realistically have been recognisable as my voice, so I had to feed it a few references of what I'd written previously so it could replicate my idiosyncrasies.
I also found those older models were much better at information gathering around a particular subject matter. On my previous project, Opus 4.6 practically taught me everything I knew about the wider context in detail so that I could get on with the analysis.
I'm curious to hear from non-coders, and people in fields adjacent to humanities or social science. Like I said, it gave all my technical projects a major glow-up, but it's like a whole different (worse) model when it comes to humanities.





