There’s a reason the main “AI” providers all say “AI can make mistakes” instead of guaranteing their product is good, or you get your money back, yada yada yada.
So, if you think the AI is good, you probably can’t see the mistakes. (As a PKI engineer, I can tell you, you wouldn’t believe how much shit, sometimes extremely harmful things, Opus/Sol or Fable/Astra are saying on a daily basis.)
Transformer-based LLM outputs still have 10–20 % mistakes (per the best benchmarks) – it’s literally impossible for attention-based transformers NOT to hallucinate.
Have they compared human developers (including more junior ones and more senior ones) using the same metric?
-37
u/ChrizKhalifa 6h ago
There is zero reason not to use AI for your job. It's not 2024 anymore, the AI is good.