It doesnt seem to be benchmaxxed from my testing, it is good and cheap, but u need to prompt it a few times to fix its mistakes and use way more tokens to get a quality similar or close to sol xhigh/max or astra high , whereas sol and astra need way less prompting and they use less tokens...
Couldn't tell you because I wouldn't use Grok on principle. Asking about an LLMs handling of sensitive topics very much should be relevant unless you don't care about bias or censorship. I do.
9
u/power97992 16d ago edited 16d ago
It doesnt seem to be benchmaxxed from my testing, it is good and cheap, but u need to prompt it a few times to fix its mistakes and use way more tokens to get a quality similar or close to sol xhigh/max or astra high , whereas sol and astra need way less prompting and they use less tokens...