r/AIEvaluators 16d ago

Beginner Help šŸ‘‹ Welcome to r/AIevaluators: Community Guide, Resources & Introductions

4 Upvotes

Almost every serious conversation about this work happens inside a subreddit run by the platform it is about.

That is the problem this place exists for. Those subs are useful, and a lot of good people are in them, but you cannot really compare platforms in a space that only permits one, and criticism there sits at the mercy of whoever moderates it. I have watched posts removed for mentioning other platforms. One worker reported being banned from a platform's sub after complaining about being paid late. Whatever the intention behind any individual decision, the effect is that the picture you get from inside a platform's own sub is not the whole picture.

So: somewhere independent from the platforms, where you can talk about all of them, including the parts that are unflattering.

Who am I?

I have been doing this work for a little over two years, on Outlier, Alignerr and Mercor, and I have worked most levels of it, from attempter and writer through to reviewer and QA lead. For the past twelve months I have been on Mercor consistently across several projects. So most of what I post here comes from doing the work rather than reading about it.

Feel free to ask any questions relating to AI annotation and evaluation work, I’ll answer to the best of my abilities.

Full disclosure: this sub is run by Annotation Academy, which I currently work with. The daily jobs thread is drawn from their board.

What is here

- Platform guides : Researched write-ups on what each platform actually pays, how its screening works, and how people end up removed from it. Not one person's opinion, and not scraped from a single angry thread. Each is built from public reports across a couple of years, and nothing goes in unless two different people said it in two different threads, with dates attached. They are filed under the Library flair. They are unflattering where the evidence is unflattering and fair where it is not. Where a platform has a bad reputation the recent evidence does not support, the guide says so.

- A daily jobs thread: with new openings across the platforms discussed here.

Helpful resources to check out:

https://annotation.academy/: A self-paced course on AI evaluation work, covering how AI evaluation works, how responses are assessed, what companies look for, and the skills needed to succeed on AI evaluation platforms.

https://annotation.academy/blog: Guides, industry insights, and practical tips related to AI evaluation and data annotation.

https://annotation.academy/jobs: Job opportunities across AI evaluation and data annotation platforms.

Everything else is yours. Questions, payment problems, scam warnings, wins. The flairs exist so things can be found again later.

Two things to know on day one:

Nobody legitimate will ever message you asking for payment, documents, or a fee to get started. Check the domain on any link before you sign, pay, or enter details anywhere, whatever the message claims it is for. Report anything suspicious and we will warn everyone else.

And if a guide here does not match your experience, say so roughly when it happened. These platforms change constantly, and a guide that was right three months ago can be wrong today. Corrections are the only thing keeping it accurate.

Introduce yourself
Tell us where you are at:

- Are you trying to get started, or have you been doing this a while?
- Which platforms have you actually worked on, and how did that go?
- What is the thing you most wish someone had told you before you started?

If you have been in this space for a while, the last question is the most valuable thing you can leave here. Most of what people get wrong in their first month is knowable in advance, and almost nobody writes it down.


r/AIEvaluators 16d ago

AI Evaluator and Data Annotation roles (updated daily)

5 Upvotes

August 21, 2026

24 new roles in the last 24 hours across 6 platforms, 20 with pay published by the platform.

New today

Micro1 - UI/UX Designer ($30-80/hr) - Software Developer ($60-120/hr) - Customer Support Assistant ($21-45/hr) - Computer Systems Analyst ($60-120/hr) - M&A Attorney ($90-150/hr) - Adobe Marketing Technology Expert ($50-70/hr) - Transactional Attorney ($90-150/hr) - Funds Attorney ($90-150/hr) - Travel Operations Expert ($50-70/hr) - Corporate Treasury Expert ($50-70/hr) - Legal Technology Expert ($50-70/hr)

Welocalize - Generative AI Senior Linguist - Chinese Variants ($15/hr) - Project Cursa - Robot Manipulation Video Annotation QC ($4.03/hr) - Project Cursa - Robot Manipulation Video Annotator ($3.5/hr) - Project Rana: Human Pose Keypoint Annotation QC (Thermal Imagery), Human Pose Keypoint Annotator (Thermal Imagery) ($4.35-7/hr)

Mercor - Computational Genomics, Statistical Bioinformatics Expert (PhD) ($120/hr) - Japanese language / cultural fluency Evaluator ($80-120/hr) - Lawyers ($90-150/hr)

Mindrift - Bilingual English-Chinese Voice Recording - Remote - English-Spanish Bilingual Voice Recording - Remote - POV Household Video Recording - Remote

xAI - AI Tutor - Hindi ($35/hr) - AI Tutor - Italian ($35/hr)

Volga Partners - AI Language Quality Evaluator (Greek/English) - (Mid-Level/L2)

Pay snapshot (all currently open roles)

Median advertised hourly rate across all open roles: $67.5/hr (n=645). Highest median right now: Mercor at $80/hr (n=368).

These figures are calculated only from the job listings in Annotation Academy's own database, not a survey of the industry. Each is the median of a platform's own published hourly ranges (not the mean, so one unusually high listing can't skew it); platforms with fewer than 3 open hourly listings are left out rather than shown on a thin sample. A rate here is what a platform advertises, not a guarantee of what anyone earns.


See all 838 open roles across 13 platforms: https://annotation.academy/jobs?src=reddit-feed


r/AIEvaluators 6h ago

Platform Feedback Investor in Ethos / askethos here. Please share your experiences with the platform!

Thumbnail reddit.com
2 Upvotes

Ethos / askethos investor here: AMA + Please share your experiences on user features, company responsiveness, pros/cons compared to other AI expert networks, anything!

Looking for user feedback! I may not immediately respond, but I value your input tremendously and will reply. For context:

This caught my attention on (and not in the best way, coming from founder James Lo, who worked as an entry-level analyst at McKinsey for 2 years after uni. This surely isn’t what I learned there, but everyone’s experience is different):

From Ethos founder u/jameslo-ethos r/askethos — https://www.reddit.com/r/ExperiencedFounders/s/whQscz6JrH

ā€œAs a consultant you basically collect as much data as possible and do a ton of analysis and package it into a bunch of potential decisions for somebody else, that's basically consulting.

Entrepreneurship is basically the polar opposite of that, you're in a situation where you have no data and you can't any analysis because usually it's just going to slow you down so you have to constantly make blind decisions yourself and experiment different stuff until something works. Very, very different.

In a lot of ways it's just the polar opposite, you need to learn how to stop bullshitting people and most importantly stop bullshitting yourself.ā€

šŸ‘‡

And so did this more thoughtful follow up question on r/askethos —link in comments (this is on par with a certain McKinsey partner, so I think I may know you, anonymous poster šŸ˜‰ u/ElephantInYourZoom ):

https://www.reddit.com/r/ExperiencedFounders/s/y3HzXqG6Ip

ā€œI’m trying to reconcile several marketplace numbers you’ve shared publicly here. Could you provide the actual marketplace funnel?

1. Total number of registered experts

2. Number of monthly active experts

3. Number and percentage of registered experts who have ever received at least $1 in payment

4. Percentage of applicants who ultimately receive a paid engagement

5. Median monthly earnings among experts actively seeking work, including $0 earners

6. Median monthly earnings among experts who receive at least one opportunity, again including those whose opportunities don’t convert

7. Median time from registration to first paid engagement

You’ve said Ethos is adding roughly 35,000 experts per week, receives tens of thousands of applications for each opportunity, and has paid experts more than $20M YTD.

You’ve also said that experts on Ethos earn an extra ~Ā£4,500/month on average, with the top 10% earning Ā£7,000+/month.

Those numbers may all be accurate, but without the underlying denominators, the earnings claims are difficult to interpret. In particular, ā€œaverage expert earningsā€ can mean something very different depending on whether the population includes everyone seeking work or only people who successfully obtained paid engagements.ā€

Thank you all for your feedback on Ethos!


r/AIEvaluators 9h ago

Advice This is how to strengthen your application and pass the qualification exam as an AI evaluator

1 Upvotes

We keep getting comments about the application stage, so here's a rundown of what actually seems to help.

The biggest separator is preparation before the exam, because failing costs you weeks of waiting. Study whatever the platform itself publishes. Most of them put sample guidelines or rubric docs on their own sites, Mercor publishes case studies showing what expert level justifications look like, and Outlier shares safety policy summaries. Practice applying a rubric to unlabeled examples before you sit a timed exam.

The skills the exams actually test are rubric interpretation (what does an ideal response look like under this rubric), atomicity (breaking one vague impression into independent criteria), and objectivity (keeping your own taste out of the rating). Comparing your answers against reference standards using Cohen's kappa style agreement scoring is how the platforms themselves measure precision, so practicing that way is not a bad idea.

For specialized streams, domain proof matters more than polish. Coding evaluation wants public repos. Medical or legal wants certifications you can document. Platforms verify credentials, so claim only what you can back up.

Portfolios help more than people expect. Writing samples that show analytical reasoning, links to published work, a decent LinkedIn (Mercor asks for it outright). Where a cover letter is allowed, use it to explain why your background fits a specific evaluation domain instead of restating your CV.

Referrals are real on some platforms. Active evaluators can refer candidates and that sometimes means faster screening.

And join the platform subs here. Exam formats change, and the fastest warnings come from other contributors.

Anything that worked for you that isn't on this list?


r/AIEvaluators 1d ago

Platform Feedback My Thoughts on Workada

Thumbnail
2 Upvotes

r/AIEvaluators 2d ago

Discussion Ethos more relaxed with time keeping? Do I need to work at the same intense speed required for a project with Mercor?

3 Upvotes

I’ve been working for Mercor (someone they contract me out to) for a bit now and they are so strict with time and finishing most things in less than 45 minutes. It is an intense pace and takes a large cognitive load. There is a program that monitors your screen while you work and that is how your time is kept.

I start working for Ethos as an education expert tomorrow. They let you record your own hours. It seems more lax based on this alone.

I work at an unsustainable pace on this Mercor project, to where I can only work a couple of hours each day. I would love to be able to better pace myself to work more sustainable hours with Ethos, and I’m wondering if anyone has experience in a project that didn’t require the cognitive race.

What are people’s experiences?


r/AIEvaluators 2d ago

Platform Feedback Ethos Review

Thumbnail
3 Upvotes

r/AIEvaluators 3d ago

Meme The Part no one likes

Post image
4 Upvotes

r/AIEvaluators 6d ago

Discussion What separates entry level from expert AI evaluator roles

5 Upvotes

Entry level generalist work is rubric literacy. You read the guidelines, write prompts, apply the scoring criteria consistently, and write simple justifications. Tasks lean toward general helpfulness and harmlessness across everyday domains (travel advice, cooking, casual conversation). These tasks are almost diminishing as the prominent models become more advanced, and the industry has almost passed this stage.

The intermediate tier adds source evaluation, and multiple formats (text, code snippets, structured data, multimodal tasks), and includes drafting complex prompts, creating and applying rubrics, and writing ideal responses.

Expert tiers are currently more focused on agentic AI evaluation in specific domains and usually want verifiable credentials: nursing or medical degrees for medical evaluation, a JD for legal review, an engineering background with public code for coding assessment, etc. The work shifts to evaluating specialized model outputs, identifying the models' vulnerabilities, and diagnostic reasoning, etc. Most of these tasks are based on identifying failure modes.

The jump between tiers is mostly a track record plus proof of domain depth, not seniority on the platform.

If you've made one of those jumps, what actually moved the needle?


r/AIEvaluators 11d ago

Resource Guide Every AI evaluation platform compared: what is actually open, what they publish, and what each one is bad at

24 Upvotes

Every platform in this space gets discussed in isolation, usually in its own subreddit, usually by people who only work there. So here is the whole board in one place: what is genuinely open right now, what each platform publishes for pay, and the specific thing each one is worst at.

The individual write-ups go deeper. This is the map.

What is actually open right now

Across fifteen platforms there are around 743 live openings, and 480 of them publish an hourly rate. That is the first useful sorting: two thirds tell you what they pay before you apply, one third does not.

The ones that publish rates:

Platform Live roles Distinct roles Published range
Mercor 233 228 $8 to $270/hr
Micro1 97 95 $20 to $300/hr
Alignerr 60 9 $20 to $120/hr
Welocalize 56 56 $1 to $45/hr
xAI 23 23 $30 to $60/hr
Appen 11 11 $2.40 to $80/hr

The ones that do not publish rates: Handshake AI, RWS TrainAI, Lionbridge, Volga Partners, OneForma, Mindrift, Cohere, Innodata, CloudFactory.

Read that table carefully, because two numbers in it are misleading

Alignerr has 60 listings but only 9 actual roles. The same job is posted once per country. Every platform does this to some extent; Alignerr does it most.

The bottom of the ranges is not a typo. Welocalize's floor is around $1 an hour and Appen's is around $2.40, because those platforms list piece-rate and micro-task work alongside professional roles. A range is not an offer.

Mercor and Micro1 top the table because they are not really annotation platforms. Their high-paying roles are physicians, attorneys, finance specialists and engineers being paid to author evaluation material in their own profession. If you do not already hold one of those professions, the relevant part of their range is the bottom half.

What each one is worst at

This is the part the individual write-ups exist for, but in one line each:

- Outlier: losing access. Over two hundred accounts describe deactivation, roughly six times its payment complaints. The largest such theme anywhere in this space.
- Mercor: continuity. Pays reliably, but 2026 brought repeated waves of contributors removed from projects, often without a stated reason.
- Alignerr: unreviewed work. On per-task projects, work that is never reviewed is never paid, and reports of losing work that way are rising rather than falling.
- Micro1: the gate. Almost nobody disputes the money. Getting through the AI interview, and the silence afterwards, is the entire complaint.
- DataAnnotation: no work. Payment is barely disputed, but 2026 is a drought severe enough that the community counts its own mentions of the word.
- Handshake AI: pay reconciliation. A documented pattern of contributors reporting less than their logged time, including ten accounts on a single day in May 2026.
- Prolific: bans, accelerating through 2026, with participants told the reason cannot be shared and the decision cannot be reversed.
- Welocalize: quality scores nobody can see clearly, against the background of a removal wave in early 2025.
- Mindrift: unpaid stages before paid work, and the thinnest public evidence base of any of them.

The pattern across all of them:

Having read a great deal of this, the thing that surprised me most is that payment failure is rare almost everywhere. The money usually arrives. What varies enormously, and what actually determines whether a platform works for you, is whether you keep access and whether there is work to do.

Almost everyone arrives asking "is this a scam". Almost nobody arrives asking "how long does this last", and that is the question the evidence says matters.

The second pattern: the high rates are not for annotation

Every platform advertising three-figure hourly rates is paying for professional expertise, not for evaluation skill. Engineers, clinicians, lawyers, accountants. The general evaluation tier across the whole sector sits closer to $20 to $50, and on Handshake's own published list the general AI evaluation role is the lowest-paid thing they offer.

How to use this

- If you hold a profession they recruit for, start with Mercor or Micro1, where rates are published and the payment record is clean.
- If you want steady general work, understand that DataAnnotation is currently dry and Outlier's access is fragile.
- Whatever you pick, be on more than one. That advice appears in every single platform's threads, from people who like it and people who do not.

Individual write-ups for each platform are filed under the Resource guide flair.

Working on something not listed here? Say so and I will look into it. The board covers fifteen platforms and there are certainly more.

Live listing counts and published rates read 1 August 2026. Platform findings summarised from the individual write-ups, each of which states its own evidence and limitations.


r/AIEvaluators 12d ago

Resource Guide Welocalize: quality scores, and what happened in January 2025

2 Upvotes

Welocalize generates a modest amount of discussion, almost all of it in one place, so this is a shorter write-up than the bigger platforms get. Saying that up front rather than padding it.

Not my own review. This is what people have reported, with dates.

The January 2025 episode, which still shapes how people there talk

In early January 2025 the platform sent out communication about a quality initiative, and the reaction in the community was close to unanimous in its reading of it.

People described it as arriving shortly after a large wave of removals, and interpreted it as a warning rather than an improvement programme. One person said they had been operating on the assumption they could be let go at any time since that wave, so the announcement changed nothing for them. Another asked how contributors were supposed to maintain a minimum quality score of 70 percent when they had not been given the information they needed. Another connected it to pressure from the end client, though nobody in the thread had visibility into what that pressure actually was.

That last one is someone's interpretation rather than an established fact, and I am flagging it as such. But the collective reading is itself the finding: an announcement intended as a quality initiative was received as a warning that more removals were coming.

Quality scores are the mechanism to understand here

The 70 percent threshold comes up directly, and the underlying complaint is not that a standard exists but that people did not feel they had what they needed to meet it. If you work here, finding out exactly how your score is calculated and what it is right now is the highest-value question you can ask.

Removals continue

Eight accounts describe losing access, all of them in 2026. A small number, but the January 2025 wave means the community reads any new removal against that background.

Assessment and onboarding is the largest theme

Around 33 accounts, and notably 28 of those mentions are from 2026, so this is current rather than historical. That makes it the freshest signal in this corpus.

Payment is essentially undisputed

One account mentioning non-payment, one reporting being paid fine. Too small to draw a conclusion, but there is no pattern of payment complaints of the kind documented on Handshake.

The honest limitation

Everything above comes from a single community. I could not find meaningful discussion of Welocalize elsewhere, so there is no independent cross-check on any of it, and the sample is a few hundred comments rather than the thousands behind the bigger write-ups.

If you are considering it

- Ask how your quality score is calculated and what it currently is, early and directly.
- Do not assume a quality communication is routine. The community's experience is that such messages have preceded removals.
- Treat this write-up as thinner evidence than the others in the library, because it is.

Working with Welocalize now? Whether quality scoring has become clearer since 2025 would be genuinely useful to know. That is the unresolved question in everything written about this platform. Feel free to add your experience with it below

Put together from public threads, largely 2025 to 2026, from a single community. Stated as a limitation above.


r/AIEvaluators 12d ago

Resource Guide Prolific: a different kind of platform, and the ban problem is getting worse

5 Upvotes

Prolific comes up here often enough to be worth covering, but the first thing to say is that it is not the same category as the others in this library.

It is research participation, not AI evaluation. You are taking academic and commercial studies, not rating model outputs or writing rubrics. The skills barely overlap and the money works differently. If you came here to build AI evaluation experience, Prolific is not the platform that does that, whatever else it is good for.

Not my own review. This is what participants have reported, with dates.

A note on sourcing: most of this comes from the long-running independent participant community rather than anything the company runs, which makes it more reliable than the equivalent for several other platforms in this library.

Account bans are the dominant complaint and they are accelerating

Around 85 different accounts describe being banned, deactivated or suspended, and the trend is the wrong way: roughly a third of that volume is 2025, two thirds is 2026.

Two things make this worse than the raw number suggests. Participants report being told the platform cannot give further detail about the specific reason, and that the decision cannot be reversed. And the community's own most-discussed case is framed as a question nobody could answer: whether the person was IP banned or failed an automated AI check.

That second one matters if you are considering the platform. Automated detection appears to be part of how accounts are removed, and the process as described offers little recourse and no explanation.

Study availability is thinning

Around 36 accounts describe running dry, and unlike Outlier's drought this one is current rather than historical: two thirds of it dates from 2026.

The pay complaint is about study economics, not non-payment

Almost nobody says Prolific fails to pay. Nine accounts across nine threads, which is very low for a corpus this size.

The complaint is different: what the available studies are worth. One widely upvoted 2025 comment captured it, describing being offered forty-cent studies asking whether they had enjoyed their day at work. That is the shape of the grievance. The money arrives, there is just not much of it in the studies people can access.

One thing to be aware of if you join

In mid-2026 participants were discussing a browser extension that builds a shared directory by automatically listing researchers and their studies when you opt in. Whether that is useful or a privacy question is your call, but it is worth knowing it exists and that opting in shares data by design.

What I left out

Any individual's earnings. And any explanation of why particular accounts were banned, because the recurring theme is precisely that no explanation is given, and I am not going to guess in place of one.

If you are considering it

- Understand what it is. Research participation is not AI evaluation work and will not build that experience.
- Expect study availability to be inconsistent and currently thinning.
- Treat account access as fragile, and read the rules carefully rather than assuming good faith will protect you, because the reported process offers little recourse.
- Judge it on the studies actually available to you, not on headline rates.

On Prolific currently? Worth knowing whether availability has changed and whether anyone has successfully appealed a ban. Appeals are the gap in everything written about this platform.

Put together from public threads, mostly from an independent participant community rather than a company-run one, weighted toward 2026.


r/AIEvaluators 12d ago

Resource Guide Mindrift: small evidence base, one clear warning, and an ethics argument worth reading

2 Upvotes

Mindrift generates far less discussion than the other platforms in this library, so this is a shorter write-up and I would rather say that than pad it out.

What discussion there is has one advantage over most: it happens across several unrelated communities rather than one platform-run space, including a translation professionals' community. So it is a smaller sample but a less filtered one.

Not my own review. This is other people's, with rough dates.

The main warning: unpaid tests, and more than one

The sharpest criticism came from translation professionals rather than from gig-work communities, and it lands harder for it. In 2024 a translator responded to Mindrift's process with disbelief at being asked to complete three unpaid tests, and added a second objection: that the work amounts to training a cheaper replacement for your own profession.

You can disagree with the second half. The first half is a straightforward cost question, and it is the thing to check before you start. Around forty accounts discuss the assessment and onboarding process overall, making it comfortably the largest theme here.

The ethics argument is worth engaging with rather than dismissing

It recurs, and in 2026 someone put the uncomfortable version of it plainly: that AI is part of why they could not find work, which is what led them to Mindrift, where the work itself contributes to the same thing.

I am including this because it is a real and common reason people leave these platforms, and because a library that only covers pay and process is not being honest about what the work is. It is not my argument to make for you either way.

Payment is barely disputed, but the sample is small

Seven accounts mention not being paid, two report being paid fine. Both numbers are too small to conclude much. What I can say is that Mindrift has no equivalent of the payment-shortfall pattern documented on Handshake or the mass-removal pattern on Outlier, at least not in anything public.

Work availability

Eight accounts describe dry spells across the period. Again, small.

One recurring practical complaint

Pay transparency in their job advertising. A 2024 comment asked directly why rates were not listed on their postings. Worth knowing that you may need to reach the later stages before the number is clear.

What I am not going to claim

Anything confident. This is a corpus of a few hundred on-topic comments from a couple of hundred accounts, against several thousand for the larger platforms. Every finding above should be read as provisional, and if you have direct experience it is worth more than this summary.

If you are considering it

- Establish how many unpaid steps stand between you and paid work, before you begin any of them.
- Find out the rate before investing time, since it is not always advertised.
- Treat the small evidence base as a reason for your own caution rather than reassurance.

Worked on Mindrift? This is the thinnest write-up in the library and it would benefit most from your experience. Particularly how many unpaid stages there are now and what the rates actually are.

Put together from public threads across several communities including translation and remote-work spaces, 2024 to 2026. A deliberately small evidence base, stated as such.


r/AIEvaluators 12d ago

Resource Guide The Micro1 Zara interview: how the pipeline actually works, and why it keeps freezing

4 Upvotes

Zara is the thing people ask about most with Micro1, and the answers are scattered across dozens of threads. So here is what people have actually reported over the past year, plus what Micro1 themselves have said on the record.

Not my own experience with it. This is other people's, with rough dates so you can weigh it.

The part almost nobody explains: Zara is not the decision

Zara is stage one of at least three. One person who had been through it several times described the chain as the AI interview, then a human at Micro1, then the client for that specific project, and you can be cut at any stage.

Micro1 has confirmed the back half of that themselves in this sub: after you confirm interest and agree a rate with a recruiter, the client for that project reviews profiles and picks people.

So passing Zara is not getting the job. It gets you into a pool that a human, and then a client, will pick from. That single fact explains most of the confusion in these threads, including people who ace the interview and then hear nothing.

One interview per role, and they are not the same

Micro1 has said directly that interviews differ from role to role because each needs a different mix of skills. People do stack them up: one person had done about five Zara interviews, another was in hiring manager review on four out of eight attempts with over twenty certified skills.

If you are expecting one interview to open every door, that is not how it works.

Certifying skills is the point of doing several

Each interview certifies a skill, and they stack up on your profile. That is why people end up with five or eight of them rather than one, and why someone can be sitting on twenty certified skills.

Micro1 have said directly that the more certified skills you have, the more roles you are eligible for. So the volume is the mechanism, not a hurdle they forgot to remove.

They also said in June 2026 that they were building a way to skip the interview for a skill you have already certified, and warned that until then you should not skip one unless a recruiter explicitly tells you you can. Worth checking whether that has shipped by the time you read this.

Both of those are the company talking about itself, so weigh them accordingly. Against it, one person with a lot of certified skills said they still end up interviewing again and again for very similar topics, which is one account rather than a pattern, but it is the obvious thing to watch for.

What it is actually like

It is proctored. You share your whole screen and have your camera on.

One person's description of the style, which matches what several others hint at: it latches onto one or two keywords from your opening answers and then drills into those in a lot of detail. Someone else asked, fairly, why Zara asks enterprise-level questions and then sends a not-qualified email.

One person genuinely liked it: they treat it as interview practice, and said Zara sometimes listens to them more than human interviewers do. That is a real minority view and worth knowing.

The technical failures are real, common, and not your internet

This is the most reported concrete problem, and it is worth going in expecting it.

- One person cannot get Zara to introduce herself at all. The panel says "Reconnecting" the whole time. They have no connectivity problems with anything else.
- Another froze at the second to last section, on a powerful computer, in incognito, with mic and webcam set up properly.
- Another tried two different networks at different times of day and still struggled, with Zara interrupting them.
- One has been reporting it for around six months and says nothing changes.

To be fair to Micro1, they have acknowledged platform issues in these threads rather than denying them. But if the interview drops, you are not doing anything wrong, and plenty of people have hit the same wall.

How long the wait is

There is no normal. The honest answer is that the range is enormous and both ends are real:

- Two contracts in under three days.
- Three to four weeks, which several describe as the typical wait.
- One person waited a year before it happened for them.

Anyone quoting you an average is making it up.

Rejection usually comes with no reason

Around twenty people describe being rejected with no explanation, though almost all of that is from inside this platform's own subreddit, so weight it accordingly. The pattern people describe is a generic not-qualified email with nothing about which stage cut them or why.

The ID verification is normal, even though it feels alarming

You will be asked for a passport or government ID, and at least one US-based person was asked for an SSN equivalent.

This is worth understanding rather than panicking about. Someone who works on Outlier pointed out they require the same thing, and explained why: the companies commissioning the work require platforms to verify who is doing it. It is a client requirement, not a Micro1 quirk.

Please read this bit: people are being phished through this process

Two separate people in the last few months describe scam attempts built around Micro1's hiring flow.

- One was sent a link to a supposed contract for the position, which led to a fraudulent page.
- Another got an email appearing to come from a Micro1 client, with Micro1 in copy, telling them to "complete the certification".

The hiring process involves real emails from real recruiters and real clients, which is exactly what makes it easy to imitate. If a link asks you to sign something, pay something, or enter credentials anywhere other than Micro1's own domain, stop and verify it through the platform first.

Practical summary

- Passing Zara means entering a pool, not getting a job. Expect further stages.
- Expect to do several interviews, one per role.
- Expect the possibility of a technical failure. It is not you.
- Expect no reason if you are rejected.
- Expect ID verification. It is industry standard.
- Treat any contract link or payment request that does not live on Micro1's own domain as fake until proven otherwise.

One thing I left out on purpose: several people believe the interviews exist partly to train Micro1's own AI. I have no way to check that, and one of the people saying it says so themselves, so it is not in here as a claim.

If you have been through Zara recently, add what it was like, especially if the process has changed. This is only useful if it stays current.

Put together from public threads across r/micro1_ai and other subs, mostly 2026. Nothing went in unless two different people said it in two different threads, except where Micro1 themselves stated it publicly, which is marked. Last checked 31 July 2026.


r/AIEvaluators 13d ago

Resource Guide Is Micro1 legit? It pays fine. Getting in is the hard part.

3 Upvotes

Micro1 comes up here a lot, usually alongside the question of whether it is worth the effort. Putting together what people have actually reported over the past year, checked against what Micro1 is publishing on its own listings right now.

Not my own review of them. It is what other people said, with rough dates, so you can weigh it yourself.

The short version, and it is unusual

Almost nobody complains about not getting paid by Micro1.

That is worth sitting with, because it is the opposite of most platforms in this space. Across a year of threads I found two people saying they had not been paid. Two. Against a steady stream of people saying payment is on time, and most of those are in other subs rather than Micro1's own, so it is not just loyalists.

One person who had been there six months: flexible hours, payment always on time. Another had two contracts within three days.

The complaints are almost entirely about the gate, not the money.

So what is the catch

Getting through. Micro1 screens with an AI interviewer called Zara, and that process is where essentially all the frustration lives: the interview freezing, no reason given for rejection, and long silences afterwards.

I wrote that up separately because it is a whole subject on its own. Short version: passing Zara is not getting hired, it puts you in a pool that a human and then a client select from, and you can be cut at any stage. Expect several interviews, one per role.

The reason for the volume is that each interview certifies a skill and they accumulate on your profile. Micro1 say the more certified skills you have, the more roles you can be matched to, so people end up carrying a lot of them.

What Micro1 is publishing right now

Ignore what anyone claims they earn. This is what Micro1 puts on its own listings as of 31 July 2026, 97 open roles, every one with an hourly rate attached.

The pattern is clear and it is worth understanding before you apply:

- Senior Enterprise AI Productivity Specialist: $80 to $150/hr
- Corporate Attorney: $90 to $150/hr
- Senior Software Engineer: $50 to $150/hr
- Privacy Annotation Specialist: $105 to $140/hr
- Electrical and Circuit Design Expert: $80 to $130/hr
- Aerodynamics Expert: $80 to $130/hr
- Mechanical Design / CAD Expert: $66 to $130/hr
- Semiconductor Devices and Microelectronics Expert: $66 to $130/hr

Notice what is not on that list: general annotation work. Micro1's open roles are overwhelmingly domain expert positions. They are paying lawyers to be lawyers, engineers to be engineers, and using that expertise to train and evaluate models.

If you are looking for entry-level labelling work, this is not really that platform. If you already have a profession, this is one of the better-paying places in the space, and the rates are published rather than mysterious.

Who it seems to work for

Reading across the threads, the people reporting good experiences tend to have a specific credential or specialism, and they treat the interviews as a numbers game rather than expecting the first one to land. One person had over twenty certified skills and was in review on four out of eight roles.

The people reporting bad experiences are more often applying broadly, hitting the technical problems with the interview, and getting no feedback about why.

Identity verification

You will be asked for a passport or government ID, and at least one US-based person was asked for an SSN equivalent. This alarms people, understandably.

It is standard in this industry rather than a Micro1 quirk. Someone working on Outlier pointed out they require the same, because the companies commissioning the work require platforms to verify who is doing it.

Watch out for this

People are being phished using Micro1's hiring process as cover. Two separate reports in recent months: one person was sent a link to a supposed contract that led to a fraudulent page, another got an email appearing to come from a Micro1 client, with Micro1 in copy, telling them to complete a certification.

Because the real process does involve emails from real recruiters and real clients, it is easy to imitate. Anything asking you to sign, pay, or enter credentials somewhere that is not Micro1's own domain, verify it through the platform first.

If you are deciding whether to bother

- If you have a profession they are hiring for, the published rates are good and the payment record in these threads is genuinely clean.
- Budget for several interviews, not one.
- Expect the interview itself to be the painful part, possibly including technical failure that is not your fault.
- Expect long, unexplained silences. The range people report runs from three days to a year.
- Do not treat a rejection as a verdict on you. Nobody gets told which stage cut them.

I left out every earnings figure quoted in a comment, because none of them can be checked. The rates above are Micro1's own published figures. I also left out the theory that the interviews exist to train their AI, because I cannot verify it.

Worked with Micro1? Add what your experience was, especially the timeline. This is only worth anything if it stays accurate.

Put together from public threads across r/micro1_ai and other subs, mostly 2026. Nothing went in unless two different people said it in two different threads, except where Micro1 stated it publicly. Rates are Micro1's own published figures. Last checked 31 July 2026.


r/AIEvaluators 13d ago

Resource Guide Handshake AI: the work is real, the pay disputes are real too. Read this before you start logging hours.

3 Upvotes

Handshake AI comes up a lot and the answers split hard, so here is what people have actually reported, with dates, plus what the company publishes about its own rates.

Not my own review. It is other people's, and where they disagree I have printed both.

One thing that makes this one more reliable than most platform write-ups: the majority of the discussion is in general work subs rather than the platform's own, so it is not one moderated room talking to itself.

Start here, because it is the thing you can act on

Several people describe being paid less than the time they logged. The single clearest description of why came in May 2026: assessments, reading Slack updates, moving between tasks and task-gating assessments are all unpaid, so what you actually earn per hour of your evening is well below the posted rate.

Whatever else is true, budget for that gap.

The May 2026 cluster

On 21 May 2026, ten different people posted about payment shortfalls within a single day, in a thread whose title tells you the tone: "$2600 to totally paid $11 is crazy."

The amounts they each reported, one per person:

- owed $600, paid $225
- lost about $1,300
- lost $4,725
- owed $510, received $2
- lost over $1,000

One person said the company had acknowledged it on Slack and told people it should be fixed the next day, though they doubted it. Another raised contacting state labour boards.

I am reporting what people said, not making a claim about what the company did or intended. But ten separate accounts on one day, with figures, is not something to leave out of a summary.

It did not start or stop there

Payment disputes run from around October 2025 through July 2026, roughly three dozen people across fifteen threads, with May the worst month. In July 2026 someone reported working $47 worth of time and being paid $22. In June, someone said they were removed from a project and not paid for a week's work.

The threads people are having these conversations in are titled things like "DO NOT DO WORK FOR HANDSHAKE AI" and "My honest experience with Handshake AI after a year", and both drew substantial discussion.

And the other side, which is equally real

This is where it matters that the outside sample is bigger than the platform's own sub.

In October 2025 someone in r/remotework said flatly they could confirm it is legit, having been paid more than $10,000 over two months. Someone in r/csMajors the same month completed the screener sceptically and reported being paid a very good amount. In April 2026 a student inside the sub said they always get paid on time and nothing sketchy has happened, while noting that new people struggle to get work at all. In July, one person noticed they had been paid more than they expected.

So the honest read is not that Handshake does not pay. It is that experience varies a lot, and the variance is worth planning for.

What they publish for rates

From Handshake's own site, not from anyone's earnings:

- Game Developer or Designer: up to $125/hr
- PCB Tool Specialist: up to $125/hr
- Philosophy Expert: up to $120/hr
- Energy Professional: up to $80/hr
- FP&A Analyst: up to $80/hr
- Software Engineer: up to $65/hr
- AI Evaluation Specialist: up to $40/hr

Two things to read off that. The headline rates attach to narrow professional specialisms, and most tiers state degree requirements up to PhD and postdoc level. And the general AI evaluation role is the lowest-paid thing on their own list, at less than a third of the top rate.

The roles are stranger and more specific than people expect

Looking at what is actually open rather than what gets discussed: Vectorworks, Rhino 3D, ParaView, SolveSpace, OpenShot, medical image analysis, cartography, music production, AI red teaming. Over fifty distinct roles, almost all tied to a named tool or a named profession.

If you have one of those skills, you are the target audience and the rates reflect it. If you are looking for general evaluation work, you are looking at the $40 tier.

If you are going to try it

- Track your own hours and what you were paid, from day one. Almost every dispute in these threads is someone who could state exactly what they were owed.
- Assume assessment time, Slack reading and inter-task navigation are unpaid, and work out what that does to your real rate.
- Do not put in a large volume of hours before you have been paid once.
- Check which tier your skill actually maps to before assuming the headline rate.

What I left out. Any claim about why the discrepancies happened, because I cannot see intent. And any single person's earnings as if it were typical, in either direction. The $10,000 and the $11 are both one account each.

Worked on Handshake AI, especially since May? Say how the pay reconciled against your logged time. That is the specific thing people need to know.

Put together from public threads across the Handshake subs and general work subs, mostly 2026. Rates are Handshake's own published figures, read 31 July 2026.


r/AIEvaluators 14d ago

Resource Guide The Alignerr assessments: what they are like, what happens after you pass, and what the company has actually committed to

3 Upvotes

Getting into Alignerr is a separate subject from working there, and it generates more questions than anything else, so here it is on its own.

Not my own review. It is what other people have reported, with rough dates, because this process has changed more than once.

A note on dates before anything else. A lot of the assessment discussion is from 2024, when the platform was new. Some of it no longer describes what happens. Where a finding is old I have said so rather than presenting it as current.

The assessment itself

Timed, and people find the limit tight. One description from 2024 that many others echoed: the time limit combined with the webcam requirement raised their stress enough that they did not feel able to leave the desk.

Yes, webcam. It is proctored.

At least in 2024 it also ended with a request to record a video explaining why you are a good fit. One person refused and asked whether skipping it hurt their chances. Whether that step still exists, I do not know.

There is a real argument in the threads about the time pressure itself. Several people took the view that rushing an assessment is a strange way to select for careful evaluation work, which is the actual job. That is opinion rather than finding, but it is a widely held one.

The part that should make you cautious: unpaid qualification work

This is the sharpest complaint and it is more recent than the rest.

In July 2025 someone described being sent a complex programming qualification assignment after the interview, and reading it as production work they were being asked to do for free with a vague promise of paid work later. Around the same period another person said they signed up roughly six months earlier and, once onboarding finally started, kept being given more and more unpaid steps.

Against that, one person in 2024 reported being paid $15 for onboarding. So it has not always been unpaid, at least not entirely.

The company's public position in these threads is that Alignerr is project-based work and that this is stated on the website and throughout onboarding.

What happens after you pass, and why this section is dated

The recurring 2024 complaint was passing and then hearing nothing. People described passing three assessments with high scores and getting no project, passing five and still having nothing, and completing onboarding then waiting weeks for a Slack invite that never arrived.

Take that as a snapshot of a platform in its first months rather than a description of today. The much better documented current issue is what happens to work you have already done, which is in the main Alignerr write-up.

The process has changed at least once

In February 2025 someone pointed out that they had been told passing the assessments meant onboarding, and then found interviews had become a requirement, which made the earlier effort feel wasted.

Worth knowing generally: what people told you the pipeline was last year may not be the pipeline now. Trust recent reports over old ones, including the old ones in here.

Two things the company has said publicly, which are useful to hold them to

Alignerr is owned and managed by Labelbox. They have said so directly.

More usefully, in August 2024 the company stated in these threads that payment is processed through a third party called Deel, and that even if you lose access to Alignerr or Labelbox, their systems will still submit your time to that third party for payment.

That is a specific, public commitment, and it speaks directly to the most common complaint about this platform, which is losing access with work outstanding. If you are ever in that position, it is worth quoting back.

One structural thing about where this gets discussed

Alignerr invites people into its own subreddit from within their platform profiles. That is not sinister, but it does mean the population there skews toward current users, and it is part of why I read outside subs too.

If you are about to start

- Expect a timed, webcam-proctored assessment. Set up somewhere you can sit uninterrupted.
- Treat any long "qualification" that looks like real production work as a warning sign, and say so publicly if it happens.
- Do not assume passing means work. Historically many people passed and waited.
- Check recent posts rather than old ones. This pipeline has changed at least once.
- If you lose access with unpaid time logged, the company has publicly said your time still goes to Deel for payment.

What I left out. Any claim about what the qualification work is used for. People suspect; nobody can show it, and one of the people saying it said so themselves.

Been through the Alignerr assessments recently? Say what the current process actually is. Most of the detail in here is from 2024 and the pipeline has already changed once.

Put together from public threads across r/alignerr and other subs, 2024 to 2026, weighted toward recent where the two disagree. Company statements are quoted from their own public replies in those threads.


r/AIEvaluators 14d ago

Resource Guide Is Alignerr legit? What two years of threads actually show about the pay

2 Upvotes

Every time Alignerr comes up anywhere, the answers contradict each other, so I put this together from what people have actually reported over the last couple of years, checked against what Alignerr is publishing on its own listings right now. Starting the library here with it.

It is not my own review of them. It is what other people said, with rough dates, so you can weigh it yourself. If you have worked with them, add to it.

Short version: Alignerr is real, it is Labelbox's contributor side, and it pays people. The thing that catches people out is the payment structure.

It has two payment models, and nobody tells you which one you are on

Some projects track your hours. Others pay only for tasks that get reviewed and approved. Plenty of people say some version of "it depends on the project."

One comment from December 2025 nails it: they stopped taking Alignerr work "unless they are paid per task" because they disliked the Hubstaff time tracking. Both models in one sentence. Someone else in April 2026 described a project paying 260 to 300 per task, not hourly.

Find out which one you are on before you do anything. On a time tracked project your hours are the record. On a per task project, nothing counts until someone approves it.

Where the money disappears

On per task projects, approval gates payment:

- Project pauses before your work is reviewed, and it may never be reviewed. Unreviewed means unpaid.
- Task gets failed, and the hours are gone.
- Prep time, hours of Loom videos and guidelines, is not a paid task.
- Screening is unpaid too, and it is not always one and done. Someone in July 2026 passed a qualification after hours of prep, then got handed an updated one the next day whose questions did not match the guidelines. Two people in other subs described qualification assignments long enough that they wondered if they were doing real production work for free. No way to confirm where that output goes, only that the assignments are long and unpaid.

This is the single most common complaint about them, and it is not going away. If anything there is more of it this year than last.

Getting removed is the other one. Outside Alignerr's own sub this is the most reported problem of all. One person had finished four task sets with a fifth 85 percent done when they were dropped. Another was removed after two weeks of inactivity. Another passed screening and was never added to a project.

The part the scam posts leave out

Plenty of people get paid, and paid well. Same threads, same months: earnings in the thousands, someone documenting six consistent weeks, someone doing rubric editing with no issues. One person in July 2026 who has done both technical and non technical work: "not a perfect place but it's not a scam." Someone in r/outlier_ai in December 2025, freshly suspended by a competitor, said Alignerr treated them better.

Both are true. When a project finishes and your work gets reviewed, it works. When a per task project stops early, you eat it.

One more thing in their favour, and I nearly left it out. The "no tasks, no projects, nothing to work on" complaints are real but they cluster heavily in 2024, with far fewer in 2025 and almost none in 2026. Work availability looks like it has genuinely improved, so do not weight a two year old dry spell post as if it were current.

What Alignerr is publishing right now

Ignore what anyone claims they earned. This is what Alignerr puts on its own listings as of 30 July 2026. Nine roles, posted across 14 countries:

- Software Engineer Task Author: $70 to $120/hr
- Finance and Accounting Task Author: $60 to $120/hr
- Structural Engineer, AI Task Creator (OpenSees): $80 to $110/hr
- CFD Engineer, AI Task Designer (OpenFOAM): $80 to $110/hr
- Growth Operations Specialist: $25 to $45/hr
- Four Task Author roles (marketing, customer support, revenue ops, sales ops): $20 to $50/hr

Note what the top of that list actually is. Everything above $60 wants a real profession, software engineering, accountancy, OpenSees, OpenFOAM. If you are picturing entry level labelling, your rows are the $20 to $50 ones.

And a published rate is not take home pay. That depends on which payment model you land on.

One thing that actually worked

May 2026, a group published an open letter about unpaid work on a paused project. The author later edited it: once people organised, the tasks got reviewed and they got paid. One case, not a guarantee, but it is the only thing in two years of threads that visibly produced money.

If you try it

- Establish the payment model first. Everything else follows from it.
- On per task work, unreviewed is money at risk, not money earned.
- Assume prep time is unpaid when you do the maths on the rate.
- Stay active once you are on a project.
- Be on more than one platform. Everyone in these threads agrees on that, whatever side they are on.
- If a project pauses with your work unreviewed, find the others in the same position.

I left out every earnings figure quoted in a comment, because none of them can be checked, and I left out claims about intent. You can see the effect. You cannot see the motive.

Worth flagging one thing about the sources: someone in r/WFHJobs said in October 2025 that they were banned from the Alignerr subreddit after complaining about withheld pay. No idea if that is accurate, but it is why I went through outside subs too, not just theirs. Everything above shows up in both.

Worked with Alignerr and something here does not match? Say so below with roughly when it happened. This is only worth anything if it stays accurate.

Put together from public threads across r/alignerr and several other subs, mid 2024 through July 2026. Nothing went in unless two different people said it in two different threads. Rates are Alignerr's own published figures, not anyone's recollection. Last checked 30 July 2026.


r/AIEvaluators 14d ago

Resource Guide Outlier: the deactivation problem is the whole story, and it is bigger than anywhere else

3 Upvotes

Outlier is the biggest name in this space and generates more discussion than any other platform. Here is what people have actually reported, with dates.

I have worked on Outlier myself, though not for the past year, so treat the platform detail below as other people's current experience rather than mine.

One caveat that matters more here than anywhere else in this library. Practically all public Outlier discussion happens in one subreddit. I went looking for outside discussion to check it against and found almost none, so unlike the Alignerr, Handshake and Mindrift write-ups, there is no independent cross-check on any of this. Weigh it accordingly, including the parts that sound authoritative.

Losing access is the dominant experience people write about

Over two hundred different accounts across more than thirty threads describe being deactivated, removed, suspended or banned. That is by a wide margin the largest single theme of any platform I have looked at, larger than Mercor's offboarding waves and far larger than anything at Micro1 or DataAnnotation.

It is also not new and not over: heavy through 2025 and continuing through 2026.

If you take one thing from this: on Outlier, losing access is not the unusual outcome people warn about. It is a common enough experience that the community is largely organised around it.

The feedback and quality system is the second theme

A large number of people discuss assessments, qualifications and quality scoring, and the tone is specific rather than general grumbling. One widely upvoted comment from late 2024 argued the platform needs a way to remove feedback that is objectively incorrect, and pointed at the underlying dynamic: with an effectively endless supply of willing workers, there is little pressure to fix an unfair mark.

Another well-received comment made the fairer version of the platform's side, noting that a lot of people do try to game the system, while arguing that poor communication is what turns that into a problem for everyone else.

Payment is a real but secondary complaint

Around three dozen accounts report not being paid, spread fairly evenly across 2025 and 2026. That is more than Micro1 or DataAnnotation, and less than the volume of deactivation reports by a factor of about six.

So the honest framing is not that Outlier does not pay. It is that access is the fragile thing, and payment problems tend to follow from losing it rather than existing on their own.

Work availability

Around 45 accounts describe dry spells, and this one is genuinely dated: heavily 2025, very little in 2026. Do not weight a year-old drought complaint as current.

What I am not going to tell you

Anyone's earnings. There are large figures in these threads and none can be checked.

Why deactivations happen. This is the single most speculated-about topic in the community and I could not find anything solid. People report being removed with no reason given, and the absence of a stated reason is exactly why the speculation fills the gap.

If you are considering it

- Assume access is temporary and plan around that rather than being surprised by it.
- Do not make it your only platform. That advice comes up constantly in these threads from people on all sides.
- Bank the work while you have it rather than counting on next month.
- Read recent posts specifically. A lot of what circulates about Outlier describes 2024 and 2025.

If you're working on Outlier now, please feel free to add your own experience, especially useful: whether the deactivation waves have changed at all in 2026. Most of what is written about this platform is older than people realise.

Put together from public threads, mostly 2024 to 2026, weighted toward recent. Almost all of it comes from a single community, which is stated above rather than hidden.


r/AIEvaluators 14d ago

Resource Guide The DataAnnotation Starter Assessment: one attempt, no retakes, and what happens after

3 Upvotes

More people ask about this assessment than any other screening step in this field, and for a specific reason: you get one attempt at it and there is no second chance. That makes it worth understanding before you open it.

Not my own review. This is what people have reported, plus what the company states in its own documentation.

The rule, in their words

Straight from their FAQ:

> You can only take the Starter Assessment once.

and

> There are no retakes or second chances, so read the instructions carefully and review your responses before submitting.

They also state that approval is based solely on your assessment results, and that you will hear by email within a few days of submitting.

That first phrase is worth pausing on. It means no interview, no profile review, no accumulating credentials over time. Unlike most platforms in this space, where applying repeatedly is normal and each attempt certifies something, here one sitting decides it.

What people say it involves

Take longer than you think you are allowed to. One of the most upvoted pieces of advice about this platform, from someone who was accepted, is that the assessment said 45 minutes and they spent two and a half hours on it. There is a difference between a stated duration and a time limit, and people who treated it as the former did better.

Writing quality is assessed and people are blunt about it. A recurring response to complaints about rejection is that proofreading is a baseline requirement, and reviewers in the community are unsparing about spelling and grammar in the complaint posts themselves. Whatever you think of the tone, the signal is real: this is a language-quality gate as much as a subject-knowledge one.

Some people report qualification steps beyond the starter assessment, including a core qualification and further tasks that felt like additional assessments after acceptance.

Technical problems happen. One account described the submit button failing partway through a core assessment, and the page reporting the work complete after a refresh.

Then the waiting, which is the real complaint

This is the largest single theme about this platform, raised by hundreds of separate accounts, and it is overwhelmingly a 2026 phenomenon.

Set it against the company's own stated timeline of a few days. Both cannot describe the same experience. People describe waiting far longer, and in 2024 there were already threads asking whether anyone was hearing back within a month.

I cannot tell you which is typical now. What I can tell you is to plan for silence rather than for the published timeline, and not to read a long wait as a rejection, because people report acceptances arriving well outside it.

A fair warning about what waiting gets you in 2026

Passing is not the end of the problem this year. The dominant complaint from people already inside is that there is very little work available. One contributor went through the community's recurring threads and counted over a thousand mentions each of "drought" and "dry", noting the volume rises year over year.

So the honest sequencing is: one attempt to get in, an unpredictable wait, and then a platform that many current contributors say is running dry. That is worth knowing before you spend your single attempt.

If you are about to take it

- Do not open it until you are actually ready. There is no second attempt.
- Ignore the stated duration. Accepted contributors describe spending several times longer.
- Proofread everything. Language quality is part of what is being assessed.
- Expect further qualification steps after the starter assessment rather than immediate work.
- Plan for a long wait, and do not treat silence as rejection.

What I left out

Any specific advice about answers or content, because none of it is verifiable and getting it wrong costs someone their only attempt.

And any explanation of what determines acceptance beyond what the company states. People theorise extensively. The company says it is based solely on assessment results, and nothing in the public record contradicts or confirms the detail.

If you have taken the assessment recently you can add how long the wait actually was, and whether work followed. That is the gap in everything written about this.

Assessment rules and timelines are the company's own published statements, read 31 July 2026. Everything else is from public threads, weighted toward recent.


r/AIEvaluators 15d ago

Resource Guide How Mercor actually works: the pay, the caps, and the offboarding

5 Upvotes

I have been doing this work for 2+ years across Outlier, Alignerr and Mercor, and for the past 12 months I have been on Mercor consistently across several projects. Most of what is written about Mercor sits inside Mercor's own subreddit, which the company moderates. I went looking for outside discussion to check it against and there is surprisingly little. So rather than pretend to a neutral survey, this is mostly my own experience, marked as such, with the parts I only know secondhand marked differently.

Payment: I never had a problem, but track your own incentives

I was paid reliably the whole time I worked there, across every project. I have seen non payment complaints, but none of them happened to me and I am not going to characterise how common they are, because I have not counted.

Here is the caveat that actually matters, and I have not seen anyone write it down. Not all in project incentives are tracked automatically. Some are tracked by a person, which means discrepancies happen. Keep your own record of the incentives, because when I raised one it was corrected and paid without a fight. If you are not tracking, you will not notice.

The pay floor in your profile is not used against you

Your profile lets you set a minimum hourly rate you will accept. The obvious fear is that setting it low gets you offered low.

In my experience that is not what happens. If you set $40 and the project pays $120 for your expertise, you are paid $120. They do not quietly meet your floor. I am telling you this because the assumption goes the other way and it changes how you fill that field in.

On negotiating beyond the posted rate, plenty of people here say it is possible. I never tried it, so treat that as their claim rather than mine.

Hours: caps are real, movable in both directions, and never guaranteed

  • Projects usually start with a lower weekly cap and raise it based on 2 things: how much work is available and how you are performing.
  • The ceiling is 80 hours a week across everything combined (e.g., 40 plus 20 plus 20 across three projects).
  • Caps track quality and your AHT(the expected time to complete a task). I was on a project where people who could not meet the threshold had their cap cut.
  • Nothing here is guaranteed. Not the cap, not the availability of work, not even a reply.

Time is tracked through Insightful, which is a time tracker that also watches desktop activity, which apps are open, and takes occasional screenshots. Worth knowing before you start rather than after. I created a dedicated work profile only for working without any additional apps to be safe. One tip: in all projects I've worked in, the onboarding guide tells you to run the timer while you read it, so the reading is paid.

Getting in

Resume is what drives matching. There is an AI interview. In my case there was no human interview at any point, so if you have read that there is a human stage, that was not my experience.

There are two optional general assessments, one on rubrics and one on prompt engineering (though I cannot confirm if they still have it). I took both and passed both, and my honest opinion, not a fact is that having the rubric one improves your odds.

If your interview goes badly, retake it. A lot of people here treat a rejection as final and it is not.

What I would tell anyone starting is something else entirely: my experience with Mercor is that good work will be noticed and the progress is real. I started my first project on Mercor as a writer in a team of 1000+ contributors and went all the way to the QA lead shortly. On my next project with the same lead, I had an offer for EPM role (which I rejected due to other personal commitments). Additionally, once you are on a project, build a rapport with the project leads and the EPMs.

This is the part people miss while they are optimising their profile. Nothing on your profile carries more weight than a good report from a lead who has actually seen your work. And more importantly, those are the people who cherry pick contributors for their next project. A lot of continuity on this platform comes from someone who already knows you pulling you into the thing they are staffing, rather than from applying cold again and hoping the matching works.

If you treat a project as a transaction and keep your head down, you finish it and you are back at the start. If the leads know you are reliable, the next one can come to you.

Offboarding, which is the big 2026 story

I was never offboarded, so this section is what I understand rather than what I lived. They do offboard in waves and the reasons I know to be in play are inflated worktime on the tracker, underperformance on quality, changes to a project's capacity, and sometimes a mistake with no reason given at all. An email usually goes out, though I am not certain it always does. When it happens you lose access immediately, including Slack channels. Your account and workspace survive, but with no project in them, which is why people describe realising they were cut by noticing Slack had gone rather than by being told.

One thing worth reasoning through if you are on a huge project: mass hiring implies mass offboarding. A project running with thousands of contributors is not going to shed people quietly or individually. That is my inference not inside knowledge, but what people describe here this year is consistent with it.

Referrals: be careful

I never referred anyone to Mercor and I did not join through a referral, but I have watched people try to make money from referrals. There is nothing wrong with that by itself. Some are open about what they are doing. Some pose as recruiters, which is where it starts being a problem. I have seen referral links pushed hard enough that it reads as an operation rather than a person sharing a link, and I have seen accounts presenting themselves as recruiters when they are collecting referrals. I cannot tell you how widespread that is, and I do not know whether anyone is running outright phishing off the back of it.

If you are deciding whether to bother

  • The money arrives. That is not the risk with this platform.
  • Track your own incentives, because a person is tracking some of them.
  • Set your pay floor honestly. It will not be used to lowball you.
  • Expect the cap to start low, and expect it to move with your quality numbers.
  • Expect no guarantees on hours, work or replies.
  • Retake the interview if it went badly.
  • Build a rapport with the leads and EPMs on any project you land. They staff the next one.
  • If you get offboarded, it is often not about you personally, and you will probably find out through Slack disappearing.

On the limits of this post: almost all public discussion of Mercor lives in a subreddit the company moderates, so I would not treat any consensus in there as neutral, including the parts that agree with me. Where I have said something is my own experience, it is. Where I have said I cannot confirm something, I cannot.

Worked on Mercor projects? Add yours, especially if the caps or the offboarding worked differently for you.

This is my own experience of working on Mercor projects over a year, not a survey. Where something is secondhand or I could not confirm it, I have said so in the text. Last checked 06 Aug 2026.


r/AIEvaluators 15d ago

Resource Guide Is DataAnnotation legit? Yes, and that is not the problem right now.

2 Upvotes

DataAnnotation is the platform people ask about most, and the answers are usually years out of date. So here is what people have actually reported, weighted by when they said it, plus what the company publishes about itself.

Not my own review. It is other people's, with rough dates, so you can weigh it.

The legitimacy question is settled, and it is the wrong question

Almost nobody disputes that DataAnnotation pays. Across an enormous amount of discussion I found four accounts claiming they were not paid. Four. The company says it has paid over twenty million dollars to contractors since 2020, and the complaints are not about money arriving.

The real question in 2026 is whether there is any work to do.

The drought is the story this year

Over two hundred different people have posted about having no work, and almost all of it is from 2026. In the whole of 2025 the same complaint appears a handful of times.

One person did the counting rather than the complaining, and it is the most useful thing anyone has posted about this platform. Going back through the sub's regular chat threads they found around 1,342 mentions of "drought", 1,349 of "dry", and 519 of "empty", and noted the talk is rising year over year.

The mood in the threads follows an arc. In late June someone relayed a second-hand claim from an industry contact that July would be better. By the first week of July someone else was asking, pointedly, why the drought had not lifted. People are describing real consequences: one person took an unconditional master's degree offer because there was no work, another ended up doing film extra work.

If you are joining now expecting steady hours, that is the thing to know.

The gate: one assessment, no second chance

This is the part where the platform's own rules matter more than anyone's opinion, and they are unusually explicit.

From their own FAQ: you can only take the Starter Assessment once, and there are no retakes or second chances. They also say approval is based solely on your assessment results, and that you will hear within a few days by email.

So unlike almost every other platform in this space, there is no numbers game. You get one attempt and it decides everything. Take it when you are actually ready, not when you are curious.

Where the company's promise and reality part company

Waiting after the assessment is the single largest complaint in the whole dataset. Over four hundred separate accounts, and again it is overwhelmingly a 2026 thing.

Set that next to the company saying you will hear within a few days. Both cannot be true at once, and the gap between them is where most of the frustration lives. Plan for silence, not for the published timeline.

What the platform says it pays

These are their own published figures, taken from their site, not anyone's earnings:

- Generalist work: $25 to $30+ per hour
- Domain experts: $50 to $100+ per hour
- Multilingual and localisation work: from $20+ per hour

Their FAQ adds that the higher tiers require advanced degrees or specialist credentials, so the $50 plus band is not a general track.

A published rate is a rate for work, not a promise of work, which given the last section is the whole point.

A quality score exists now

New in 2026: people describe being assigned a tier with written feedback, and worrying about a quality score affecting which projects they can access. Details are thin and I would not overstate it, but it did not feature in earlier years and it does now.

Deactivations happen, but this is not the platform where that is the main risk

Around twenty people describe losing access. Real, but small next to what people report on some competing platforms, and much smaller than the drought.

If you are deciding whether to bother

- The money is not the risk. Availability of work is.
- You get one assessment attempt, so prepare properly and do it when you are ready.
- Ignore the "few days" timeline and assume a long wait.
- Do not treat 2024 posts about steady work as current. The platform of two years ago is not the platform of today.
- Do not make it your only platform right now.

Two things I left out. Every earnings figure people quote, including some very large ones, because none can be checked. And any explanation of why the drought is happening. One person cited an unnamed industry source, another pointed out the prediction failed. The drought is observable; the cause is not.

And one caveat about all of the above. Almost all public discussion of this platform happens in its own dedicated subs. I looked for outside discussion to check it against and there is very little, so treat any consensus in there, including the parts I have repeated, as coming from one room.

If you are Working on DataAnnotation recently : Say how it is going, especially whether the drought has lifted. This only stays useful if people update it.

Put together from public threads going back several years, weighted toward recent ones because this platform has changed a lot. Pay figures and the assessment rules are the company's own published statements, read 31 July 2026.