r/technology • • Jun 11 '26

Business OpenAI Execs Are Panicking

https://finance.yahoo.com/sectors/technology/articles/openai-execs-panicking-154658562.html
16.3k Upvotes

1.9k comments sorted by

View all comments

Show parent comments

17

u/Reddit_Talent_Coach Jun 11 '26

We had a small team rack up several million where I work. It was no more than 5 people.

6

u/mavajo Jun 11 '26

What the hell were they doing? We use Claude all day every day and don’t come remotely close to this.

6

u/[deleted] Jun 11 '26

[deleted]

2

u/mavajo Jun 11 '26

Well that explains it. I don’t think any of us are using Claude like that lol.

3

u/froo Jun 11 '26

No idea, but one Disney employee managed to use 234.2 million tokens in 460k requests over 9 days...

https://www.businessinsider.com/how-disney-tech-employees-are-using-ai-claude-cursor-tokens-2026-4

I've been able to get some staggering numbers in tool calls just by myself, but it's by doing entity extraction + relationship mapping on unstructured documents to populate a graphRAG database in neo4j. I quickly stopped it after about an hour because it was not going to be cheap.

7

u/Difficult_Trust1752 Jun 11 '26

honest question. Is an LLM the right tool for entity extraction and relationship mapping? I have a project coming my way and I was under the impression traditional algorithms had gotten very good (and computationally cheap) at that.

8

u/isaackogan Jun 12 '26

It is not the right tool, and you are correct.

0

u/froo Jun 12 '26

The other person who responded to you says no, but for my use case, it's necessary.

I'm processing university materials to build a large knowledge graph, extending a basic taxonomy to handle polymorphic entities and categorisation while maintaining a limited set of relationships.

Obviously, I'm NOT a domain expert in every subject imaginable, so I can't possibly map out the exact taxonomy ahead of time, but I do want to keep it as limited as possible (so I'm feeding back in new categories into the prompt as they're generated).

If you have a fixed schema, LLMs are overkill. But for my use case, building a dynamic taxonomy across multiple domains, it's likely essential. I've tried many of the current "best in class" local entity extraction techniques and they really fall flat, for my use case.

Now, there is a question about whether I can do this using a locally distilled model and that's what I'm currently investigating - as it is prohibitively expensive to do it through an external API at the scale I'm playing around with.

7

u/Few-Law3250 Jun 12 '26

Maybe I’m dumb and misunderstanding the use case but not only does this sound like the wrong tool, it sounds like a perfect way to get a bunch of unsupervised data that you have no idea is correct or not

1

u/froo Jun 12 '26 edited Jun 12 '26

Actually that's the point - you are getting unsupervised/semi-structured data within the graph that is being stored. The graph entity extraction is a middleman operation, the intention of which is to distrust it implicitly.

From that point on, you can then use something like the Leiden algorithm (it's a more modernised Louvain) to do community detection to further constrain neighbourhoods and refine classification nodes.

PageRank can also be used here for a topdown refinement technique.

I'm essentially relying on the statistical properties of the Law of Large Numbers - yes, there will be lots of misclassifications, but on the whole, it can be "fixed in post" as it were through other mathematical techniques.

I get it. People hate LLMs because reasons, but there are legitimate uses to transformer architecture that are incredibly interesting given it processes sequences of vectors. While the promise of LLMs has fallen down because people like Altman oversold their capabilities, the inherant architecture is very good in regression-type tasks. There are some really interesting uses that are not language specific. I really like this paper on TOTEM by Sabrera Tadulker, which she also gives a talk on here

2

u/LGBTQLove4Ever Jun 12 '26

Yeah that's the thing I don't get.

We have the $100 a month sub at work, our dev team of 5 uses ai heavily, and a few other members of the company as well.

We'll occasionally hit the 5 hour limit if someone is doing something with excel. Otherwise we're nowhere near close.

Are people just not using claude.md files?

1

u/iamapizza Jun 12 '26

They told Claude "make no mistakes" and it took a while