Overview
This is a short-term human-AI interaction evaluation pilot assessing whether providing social context before a user request helps a language model generate more relevant and personalized advice. Contributors create realistic social scenarios, interact with the model under different conditions, and evaluate response quality.
What You’ll Do
- Create realistic and ongoing social scenarios involving personal, professional, or interpersonal context
- Interact with the language model under assigned evaluation conditions
- Evaluate model responses for relevance, usefulness, personalization, clarity, and alignment with provided context
- Explain the reasoning behind each rating through clear and detailed written feedback
- Follow project guidelines and complete assigned tasks within the required timeline
Requirements
- Strong written English proficiency and clear communication skills
- Creative and analytical thinking for developing realistic social scenarios
- Educational or professional background in Business Analysis, English, Literature, Journalism, Creative Writing, Scriptwriting, Content Writing, Film and Video Captioning, or related field
- Ability to understand nuanced interpersonal situations and evaluate AI-generated responses objectively
- Three or more years of experience in data annotation, human-AI interaction, content evaluation, or similar analytical projects (preferred but not mandatory)
Who Should Apply
Apply if you have strong writing skills and a background in English, creative writing, journalism, or business analysis—you'll thrive creating nuanced scenarios and evaluating AI output with precision. Skip this if you're uncomfortable with ambiguity in evaluation criteria, prefer longer-term stability, or lack experience explaining subjective judgments in detail.
Salary Insight
No pay rate is stated in the posting. This is a pay-per-task engagement, which means you'll be paid for each completed task rather than an hourly or annual rate. Ask during recruitment or in initial correspondence: what is the per-task pay, and roughly how many tasks can you complete per day? With ~60 minutes per task and full-time engagement, you should clarify the total earning potential over the 2-3 week project.
About Turing
Remote work platform connecting skilled professionals with AI labs and companies for software engineering, AI training, data, and specialist contract work.
Remote policy: Remote worldwide, but location eligibility varies by individual project. Many AI roles require minimum hours per week and U.S. Pacific Time overlap. Contractors and freelancers only, not employees.
How hiring works at Turing →Required Skills
Compensation
$7/per task
- Location
- Worldwide
- Engagement
- Contract
- Posted
- Oct 4
Opens Turing’s listing on MyRemoteJobs — we don’t collect applications ourselves.
About Turing
Remote work platform connecting skilled professionals with AI labs and companies for software engineering, AI training, data, and specialist contract work.
Remote policy: Remote worldwide, but location eligibility varies by individual project. Many AI roles require minimum hours per week and U.S. Pacific Time overlap. Contractors and freelancers only, not employees.
How hiring works at Turing →Will your CV get past the filter?
Most applications are rejected by software before a person ever reads them. Scan your CV using our ATS Checker tool, see exactly what's blocking it, then fix it in the CV Builder and apply knowing it'll get through.
Scan my CVGet roles like this by email
Tell us what you do and we'll email you when a similar remote job is posted. No account needed.
See our jobs more often in your search
Add MyRemoteJobs as a Preferred Source on Google. Free, takes seconds, and you stay on this page.
Application Tip
Lead with a specific example of scenario creation or evaluation work you've done before—show that you can write realistic, layered social situations and critique responses with clear reasoning, not just a skills list. The per-task structure and tight 2-3 week timeline mean consistency and speed matter as much as quality; mention any experience meeting tight deadlines in similar evaluation or annotation work.
Sourced from MyRemoteJobs · verified against the original posting
Similar open positions
Quantitative Rates Researcher
micro1
VerifiedVideo Data Annotator
micro1
VerifiedVideo Data Annotation Specialist
Turing
Verified