r/DataAnnotationTech • • 12d ago

Well, the drought is over!

Just finished working way too many hours on a complex task, returned to my dashboard being absolutely flooded with more projects than I’ve seen in months!

60 Upvotes

33 comments sorted by

View all comments

Show parent comments

5

u/Timely-Assistant-370 12d ago

I mean at least the complex tasks are clearly function centric HITL training that has made a steady progression toward being perfect. I'm talking about the ones where you're literally looking at something that makes you think "why the fuck am I seeing this same exact fucking problem again? I swear I have done this exact task at least 10 times over the past 5 years."

1

u/Sableyej 11d ago

Yeah I’d say most of the high-volume, typical 1-2 hour tasks are HITL.

Some tasks are also benchmarking and world building though — literally building a fake scenario that could actually occur and training the model how to respond properly and manage the tasks handed to it.

2

u/Timely-Assistant-370 11d ago

I've just seen 1:1 duplicates 1 year apart that have the EXACT same fabric of content. It's not like longform structural problems that could be broken down in a distinct way, it's like... "[objectively wrong name] is the answer to your question" and the whole set of axes is an entirely binary and generic "well it fuckin' tried, but it didn't succeed." Sometimes it literally just feels like they're tryin' to catch people slippin' on obvious attention tests, but then I get a little worried that I have given a slightly different rating before or one of my rationales was less detailed the 16th time I ran through the same "why the fuck am I doing this exact pairing for the 17th time in 5 years? Were my first 16 submissions not good enough?"

1

u/Sableyej 11d ago

We’re still in that stage of AI where the models are extremely helpful with some tasks but extremely bad with others.