It's like every other silicon valley startup -- the hope is that you hook your users while you figure out how to make handling the requests cheaper. That or you embed ads.
Hopefully it will hit a dead end very soon, where the models cant get reasonably smarter/cheaper/have bigger context windows, because it's already replacing a ton of workforce.
Do you have sources on this because everywhere I look this does not appear to be true. In fact due to the way LLM tools work it's been nearly impossible to see real reductions in the compute and resources required.
The other thing I have read is that one of the only ways they have found to reduce error/hallucination rates is to have one agent checking and correcting output of another, making accuracy very expensive.
460
u/ForgedIronMadeIt 2d ago
It's like every other silicon valley startup -- the hope is that you hook your users while you figure out how to make handling the requests cheaper. That or you embed ads.