Overview
micro1 is building evaluation services for real-world AI agents. This role involves testing two AI agents side by side as they attempt everyday errands, scoring their performance, and providing feedback to inform product improvements.
What You’ll Do
- Run real-life scenarios with AI agents (bills, returns, email follow-ups, bookings, shopping, groceries, travel, family logistics, outings, forms)
- Approve or stop AI agents at key steps during task execution
- Screen record every attempt
- Score each run on completion, quality, control issues, satisfaction, and time
- Share clear written feedback on agent experience
Requirements
- US-based
- Bachelor's degree
- Already uses AI in daily life
- 2-3 years experience preferably in engineering, business management, or office/operations
- Detail-oriented
- Strong ability to follow instructions and document each step
- Clear written communication skills
- Active personal AI account (Free, Go, or Plus tier)
- Smartphone capable of screen recording
Who Should Apply
This suits people who are already comfortable with AI tools, detail-oriented, and want flexible work on their own schedule. You'll thrive if you're good at spotting where systems fail and can articulate what went wrong clearly. It's frustrating if you prefer deep expertise in one domain or want a structured team environment — this is solo contract work evaluating systems that will often disappoint.
Salary Insight
The $25–$45/hour range is wide, likely reflecting different performance levels or experience tiers within the cohort. For contract QA testing work with minimal setup friction, this sits at market rate. The posting doesn't specify how the band is determined; clarify in your interview whether movement within it depends on your background, speed, or quality of feedback.
About Micro1
Platform connecting domain experts with AI training and evaluation projects, paid on a flexible contract basis.
Remote policy: Remote worldwide, excluding Afghanistan, Belarus, China, Cuba, Democratic Republic of the Congo, Hong Kong, Iran, Iraq, Libya, Macao, Myanmar, North Korea, Russia, Somalia, South Sudan, Sudan, Syria, Venezuela, Ukraine and Yemen. Individual projects may have…
How hiring works at Micro1 →Required Skills
Compensation
$25 - $45/hour
- Location
- U.S. based
- Engagement
- Contract
- Posted
- 2d ago
Opens micro1’s listing on MyRemoteJobs — we don’t collect applications ourselves.
About Micro1
Platform connecting domain experts with AI training and evaluation projects, paid on a flexible contract basis.
Remote policy: Remote worldwide, excluding Afghanistan, Belarus, China, Cuba, Democratic Republic of the Congo, Hong Kong, Iran, Iraq, Libya, Macao, Myanmar, North Korea, Russia, Somalia, South Sudan, Sudan, Syria, Venezuela, Ukraine and Yemen. Individual projects may have…
How hiring works at Micro1 →Will your CV get past the filter?
Most applications are rejected by software before a person ever reads them. Scan your CV using our ATS Checker tool, see exactly what's blocking it, then fix it in the CV Builder and apply knowing it'll get through.
Scan my CVGet roles like this by email
Tell us what you do and we'll email you when a similar remote job is posted. No account needed.
See our jobs more often in your search
Add MyRemoteJobs as a Preferred Source on Google. Free, takes seconds, and you stay on this page.
Application Tip
Lead with a specific example of where you've caught a bug or spotted a failure in a system you use regularly — something that shows you notice edge cases and can explain *why* something broke. This role lives on that skill. Mention any experience documenting processes or giving feedback on tools, not just using them. The company will want evidence you can write clear, actionable feedback, not just "this didn't work."
Sourced from MyRemoteJobs · verified against the original posting
Similar open positions
Language Specialist – AI Language Evaluation
Meridial
VerifiedEveryday AI Agent Tester
micro1
VerifiedVideo Data QA Specialist
Turing
Verified