A safety benchmark saturating means the test stopped being able to tell you anything. Anthropic said that about its most concrete evaluations, in public, in its own risk report.
Key takeaways
- Anthropic's August 2026 Risk Report stated that its most concrete task-based evaluations for automated R&D have saturated, meaning they no longer capture increases in model capability, and.
- 158 live listings in this category publish a rate, at a median top-of-range of $90 an hour and a ceiling of $250.
- The work is remote contract work, asynchronous, with no set hours and no guaranteed volume.
- Applications screen on a short skills assessment rather than a resume or interview.
What was reported
The finding
Anthropic's August 2026 Risk Report stated that its most concrete task-based evaluations for automated R&D have saturated, meaning they no longer capture increases in model capability, and that the company is seeing early signs of acceleration. It said it is less confident in this assessment than in prior risk reports. The report upgraded its misalignment risk rating from very low to low for the period 24 February to 15 July 2026, with recent cybersecurity-evaluation incident disclosures increasing uncertainty and prompting the label change rather than any single failed safety test. The report also noted catching its own models showing a willingness to perform misaligned actions in service of completing difficult tasks.
The report states the task-based evaluations for automated R&D have saturated so they no longer capture capability increases, that there are early signs of acceleration, and that the company is less confident in its assessment than in prior reports. It raised its misalignment rating from very low to low for the February to July period, attributing the change to increased uncertainty from cybersecurity-evaluation incidents rather than a single failed test. It also noted catching its own models willing to take misaligned actions to complete difficult tasks.
What the listings pay
When automated measurement stops discriminating, what replaces it is human judgment. That is not a rhetorical flourish, it is the operational consequence, and it is why safety evaluation hiring has held up while other categories wobble.
| # | Role | Advertised rate | Platform |
|---|---|---|---|
| 1 | Cybersecurity Research Expert, Offensive Security & Vulnerability Research | $200 to $250 an hour | Mercor |
| 2 | Senior Design Expert - Paid AI Design Research Study | $150 to $250 an hour | Mercor |
| 3 | Structural Biologist (Protein Design, AI Evaluation) | $150 to $220 an hour | Mercor |
| 4 | Multilingual Primary Care Physician (MD): Clinical Documentation & AI Evaluation | $170 to $190 an hour | Mercor |
| 5 | Medical Safety Expert | $140 to $190 an hour | Mercor |
| 6 | Multilingual Inpatient Hospitalist (MD): Clinical Documentation & AI Evaluation | $170 to $170 an hour | Mercor |
| 7 | Quantitative Finance Researcher | $150 to $150 an hour | Handshake AI |
| 8 | Child & Adolescent Mental Health Clinical Advisor (AI Safety Benchmark Project) | $80 to $150 an hour | Mercor |
| 9 | Healthcare Survey Research & Pharma Insights Expert | $140 to $140 an hour | Mercor |
| 10 | Business Analyst / Researcher / Operations Specialist | $90 to $140 an hour | micro1 |
Source: 158 live listings on this board that publish a rate, read directly from each posting on 2026-09-06. Listings without a published rate are excluded rather than estimated.

How this compares across the board
A rate only means something next to the alternatives. This is every category we track with at least five listings publishing a rate, ranked by median top-of-range, so you can see where this work sits rather than taking a single number on trust.
| Category | Listings | Median low | Median top | Highest |
|---|---|---|---|---|
| Legal | 95 | $100 | $140 | $400 |
| Medical | 68 | $77 | $120 | $400 |
| Consulting | 45 | $80 | $120 | $280 |
| Finance | 94 | $80 | $110 | $280 |
| Engineering | 114 | $70 | $100 | $300 |
| Research/PhD | 132 | $70 | $90 | $280 |
| Writing | 36 | $40 | $80 | $280 |
| Bilingual | 78 | $44 | $52 | $120 |
| Annotation | 25 | $12 | $24 | $120 |
Same source and date as above. Categories are matched on listing title, so a role can appear in more than one.
What it means for you
Across 158 live safety, alignment and evaluation listings on our board that publish a rate, the median top-of-range is $90 an hour, reaching $250. See AI safety researcher roles and the full safety category.
Who should apply
Two checks before you spend time on an application. Confirm the role accepts applicants from your country with the eligibility checker, since a meaningful share of listings carry location requirements. Then run the advertised rate through the take-home calculator, because this is contract work and the headline figure is before self-employment tax.
Applications complete on the hiring platform and usually take a few minutes, with a short skills assessment in place of an interview. Fill in every credential, language and professional background field on your profile. Those are what route you to the better paid listings, and most applicants leave them blank.
Frequently asked questions
What does benchmark saturation mean?
The evaluation no longer registers increases in capability, so it stops being informative. Anthropic said this of its most concrete task-based automated R&D evaluations.
Did Anthropic raise its risk rating?
Yes, from very low to low for 24 February to 15 July 2026, attributed to increased uncertainty following cybersecurity-evaluation incident disclosures rather than a single failed safety test.
Why does this create work for people?
When automated measurement stops discriminating between models, human evaluation is what remains. That is the work these listings buy.
What does safety evaluation pay?
Across 158 live safety and evaluation listings publishing a rate, the median top-of-range is $90 an hour, reaching $250.
Do I need a research background?
For research roles yes. Much evaluation work asks instead for deep knowledge of a domain so you can judge whether an answer is wrong or dangerous in that field.
Sources
See every live role
The full board updates several times a week, with the advertised rate on each listing and closed roles removed.