Overview
micro1 is testing how well AI agents handle real-life tasks like bill payments, travel bookings, and shopping. This contract role involves testing two AI agents side by side on your own phone, approving key steps, recording sessions, and providing detailed feedback on their performance.
What You’ll Do
- Run real-life scenarios including bills, returns, email follow-ups, bookings, shopping, groceries, travel, family logistics, and forms
- Test two AI agents side by side using your own accounts on your phone
- Approve or stop the agent at key decision points
- Screen record every attempt and run
- Score each attempt on completion, quality, control issues, satisfaction, and time
- Share clear written feedback on the agent experience and what worked or didn't
Requirements
- US-based
- Bachelor's degree
- Active personal AI account (Free, Go, or Plus tier)
- Smartphone capable of screen recording
- 2-3 years of experience preferably in engineering, business management, or office/operations
- Detail-oriented with strong attention to detail
- Excellent at following instructions and documenting steps
- Clear written communication skills
- Already uses AI in daily life
Who Should Apply
Apply if you're detail-oriented, comfortable with AI tools, and can articulate clearly what works and what doesn't. This suits people in operations, business, or technical backgrounds who notice gaps in workflows. It will frustrate anyone who finds repetitive task documentation tedious, or who isn't comfortable submitting honest critical feedback on a product in development.
Salary Insight
$25–$45/hour is a wide band, likely reflecting experience level and how thoroughly you document and score each run. No indication whether this shifts based on the complexity of tasks assigned or your performance. At ~1.5 hours per run, a typical task would pay $37.50–$67.50; confirm during onboarding how many runs per week are expected and whether the rate is fixed or negotiated.
About Micro1
Platform connecting domain experts with AI training and evaluation projects, paid on a flexible contract basis.
Remote policy: Remote worldwide, excluding Afghanistan, Belarus, China, Cuba, Democratic Republic of the Congo, Hong Kong, Iran, Iraq, Libya, Macao, Myanmar, North Korea, Russia, Somalia, South Sudan, Sudan, Syria, Venezuela, Ukraine and Yemen. Individual projects may have…
How hiring works at Micro1 →Required Skills
Compensation
$25 - $45/hr
- Location
- USA only
- Engagement
- Contract
- Posted
- 2d ago
Opens micro1’s listing on MyRemoteJobs — we don’t collect applications ourselves.
About Micro1
Platform connecting domain experts with AI training and evaluation projects, paid on a flexible contract basis.
Remote policy: Remote worldwide, excluding Afghanistan, Belarus, China, Cuba, Democratic Republic of the Congo, Hong Kong, Iran, Iraq, Libya, Macao, Myanmar, North Korea, Russia, Somalia, South Sudan, Sudan, Syria, Venezuela, Ukraine and Yemen. Individual projects may have…
How hiring works at Micro1 →Will your CV get past the filter?
Most applications are rejected by software before a person ever reads them. Scan your CV using our ATS Checker tool, see exactly what's blocking it, then fix it in the CV Builder and apply knowing it'll get through.
Scan my CVGet roles like this by email
Tell us what you do and we'll email you when a similar remote job is posted. No account needed.
See our jobs more often in your search
Add MyRemoteJobs as a Preferred Source on Google. Free, takes seconds, and you stay on this page.
Application Tip
Lead with a specific example of how you use AI in your daily work or life — not just "I use ChatGPT" but "I use it to X, and I noticed Y gap." The hiring team is looking for someone who thinks critically about user experience, not just someone who can follow a checklist. Your screening answers and interview should demonstrate that you notice what broke, why it matters, and can explain it clearly in writing.
Sourced from MyRemoteJobs · verified against the original posting
Similar open positions
Language Specialist – AI Language Evaluation
Meridial
VerifiedAI Agent Evaluator
micro1
VerifiedVideo Data QA Specialist
Turing
Verified