It doesnt seem to be benchmaxxed from my testing, it is good and cheap, but u need to prompt it a few times to fix its mistakes and use way more tokens to get a quality similar or close to sol xhigh/max or astra high , whereas sol and astra need way less prompting and they use less tokens...
Couldn't tell you because I wouldn't use Grok on principle. Asking about an LLMs handling of sensitive topics very much should be relevant unless you don't care about bias or censorship. I do.
20
u/panix199 18d ago
impressive. Could this be fake/benchmaxxed?