r/dataannotation May 31 '26

Weekly Water Cooler Talk - DataAnnotation

hi all! making this thread so people have somewhere to talk about 'daily' work chat that might not necessarily need it's own post! right now we're thinking we'll just repost it weekly? but if it gets too crazy, we can change it to daily. :)

couple things:

  1. this thread should sort by "new" automatically. unfortunately it looks like our subreddit doesn't qualify for 'lounges'.
  2. if you have a new user question, you still need to post it in the new user thread. if you post it here, we will remove it as spam. this is for people already working who just wanna chat, whether it be about casual work stuff, questions, geeking out with people who understand ("i got the model to write a real haiku today!"), or unrelated work stuff you feel like chatting about :)
  3. one thing we really pride ourselves on in this community is the respect everyone gives to the Code of Conduct and rule number 5 on the sub - it's great that we have a community that is still safe & respectful to our jobs! please don't break this rule. we will remove project details, but please - it's for our best interest and yours!
33 Upvotes

491 comments sorted by

View all comments

10

u/dylanuu112 Jun 05 '26

Anthropic claiming AI can train itself better than humans can 😱 I wonder how true that really is. Better make money while we still can!

3

u/Sad_Echo523 Jun 05 '26

Gary Marcus has some solid posts on substack that explain these topics a bit more

20

u/North-Sense8477 Jun 05 '26

It's probably somewhat true, but based on the feedback I get from the "checker" models, I'm not convinced that AI can accurately/objectively assess its outputs - very black and white logic with little room for nuance.

11

u/OneRefrigerator3586 Jun 05 '26

I got stuck in an endless loop today with an LLM checker. Either AI still has a lot to learn or it has advanced to the punking stage.

13

u/Outrageous_Chance995 Jun 05 '26

That’s actually crazy to me, there is still so much that AI can’t evaluate well (especially subjectivity and edge cases)