Turing

AGI infrastructure marketplace for AI training work

Turing runs an open board of contract listings behind the AI labs it serves — software and infrastructure engineering, PhD-level science and math, legal and compliance review, finance, linguistics, and dozens of narrower domain-expert and generalist roles. Most listings are pay-per-task or hourly and state a rate; many also carry a referral bonus paid on top of the role's own pay.

Turing at a glance

Valuation
$2.2B
Series E, announced 6 March 2025 — doubled from the prior round
Series E
$111M
Led by Khazanah Nasional Berhad (Malaysia's sovereign wealth fund), March 2025
Team size
~7,000
Reported headcount, mid-2026
Founded2018
HQSan Francisco
BackersKhazanah Nasional Berhad, WestBridge Capital, Sozo Ventures

Figures reported by Turing and press coverage, not measured by Sidequest. Funding from Turing's own Series E announcement and matching press coverage, March 2025 — the most recent round found as of 15 September 2026.

How screening works

Create a free account at work.turing.com and apply directly to a specific listing. Most roles combine a short skills assessment with an async or live interview; the exact bar varies by listing rather than following one fixed funnel.

Who fits best

Software and infrastructure engineers, PhD/Master's-level scientists and mathematicians, lawyers and compliance reviewers, finance professionals, bilingual analysts, and domain SMEs across fields from healthcare to personal finance to sports — the board runs wide rather than narrow.

Open Turing roles

All roles →
New
TuringLAW

Legal Expert - Legal, Risk & Compliance - Lawyers Knowledge

Turing is hiring practicing US-admitted lawyers to write and grade expert-level legal tasks across contract drafting, litigation support, and legal research for frontier AI model training. Work is pay-per-task at an observed $400 per approved task, with senior review and calibration feedback built into every submission.

Rate as listed$400/taskView
New
TuringSTEM AND CODING

LLM Trainer - Agent Function call

Write synthetic multi-turn dialogues in which a user and a tool-using AI assistant work through tasks in apps like calendar, email, maps and drive — playing both sides, including the assistant's function calls and simulated tool responses. Turing runs this for a frontier lab; the work is contractor, fully remote, 20+ hours a week with four hours overlapping PST, on a 12-week contract.

See briefView
New
TuringSTEM AND CODING

Principal GenAI Engineer

Turing is hiring a staff/principal-level engineer to design and ship production agentic AI systems — tool-calling agents, RAG pipelines, guardrails, and evaluation harnesses — for Fortune 500 clients. It is a full-time, hands-on role based in Bengaluru with three days a week in office and a daily overlap into US Eastern morning hours.

See briefView
New
TuringGENERALIST

Finance Expert (Omnichannel Chat Support) - USA only

Turing is hiring US-based accountants and bookkeepers to answer SMB owners' financial and bookkeeping questions through omnichannel chat, including complex, data-heavy accounting queries. The engagement runs roughly four months as a contractor, 4–40 hours per week with four hours overlapping PST, and pay is not disclosed on the listing.

See briefView
New
TuringLAW

Legal Expert - Legal, Risk & Compliance - Lawyers Knowledge

Turing is hiring US-admitted lawyers to write and review expert-level legal tasks across contract drafting, litigation support, and legal research for frontier AI lab training data. Work is pay-per-approved-task at $400 per task, with an average handling time of 10–12 hours, under senior review and a fixed rubric.

Rate as listed$400/hrView
New
TuringLAW

Team Lead - Legal, Risk & Compliance - Lawyers Knowledge

Turing is hiring a senior lawyer (12–15 years' practice) to manage Pod Leaders producing expert-level legal tasks and evaluations for frontier AI labs. It is a six-week contract on PST hours, with quality assurance over contract drafting, legal analysis, and regulatory interpretation as the core of the work.

See briefView
New
TuringLAW

Pod Lead - Legal, Risk & Compliance - Legal Knowledge

Turing is hiring senior US-qualified lawyers to lead pods of 16–20 legal trainers producing expert-level tasks and gold-standard solutions for frontier AI labs. The role is half hands-on legal authoring and half quality assurance, rubric calibration, and coaching, on a six-week PST-aligned contract with possible extension.

See briefView
New
TuringSTEM AND CODING

Software Engineer Mining

Turing's Mining team turns real enterprise workflows into long-horizon tasks and grading rubrics that AI agents are trained and evaluated against. You'll author and QA those tasks as a backend Python engineer, validating them end-to-end inside Python-based replicas of tools like Slack, Jira, and Notion.

See briefView
New
TuringSTEM AND CODING

Software Engineer - QC

Turing is staffing a QC team for a frontier agent-training data initiative, reviewing long-horizon agent tasks and rubrics authored by other engineers before they ship. You reproduce the tasks inside their connector environments, debug what breaks, and write precise findings back to task authors — contractor assignment, roughly 35 weeks, 20+ hours a week with 4 hours of PST overlap.

See briefView
New
TuringSTEM AND CODING

Software Engineer with Python & Docker experience

Turing is staffing a four-week, pay-per-task contract building Python backend connectors that emulate SaaS tools (Slack, Jira, Notion, Gmail) or authoring long-horizon agent tasks with evaluation rubrics. Observed pay is $300 per approved task, with a minimum of one approved task per day and no upper cap.

Rate as listed$300/taskView
New
TuringCREATIVE

Professional Writing Human Data Collection

Turing is recruiting senior marketing practitioners to write and critique realistic professional-writing tasks — campaign briefs, case studies, corporate communications, career documents — that are used as training and evaluation data for frontier models. The engagement is a short, full-time contractor sprint: 40 hours a week with 8 hours overlapping PST, running up to about two weeks.

See briefView
New
TuringSTEM AND CODING

Software Engineer Pod Lead

Build Python backend clones of enterprise SaaS tools (Slack, Jira, Notion, Gmail) that serve as sandboxes for AI agents, then author and QA long-horizon agent tasks with precise grading rubrics. Full-time contractor assignment, 40 hours a week with six hours of PST overlap, initially five weeks and restricted to a specific list of countries.

See briefView
New
TuringSTEM AND CODING

Electrical Engineering

Turing is hiring contract engineers to author and score aerospace and flight-dynamics tasks used to train frontier AI models, building golden geometries and mission plans in OpenVSP, XFLR5, FlightGear+JSBSim, OpenRocket and QGroundControl. The posting is filed under Electrical Engineering but the stated scope is aero/flight dynamics — expect the screen to test that domain specifically.

See briefView
New
TuringSCIENCES

Scientific Coding - STEM and Python

Turing's SciCode project, run with NVIDIA, pays STEM graduate-degree holders to author multi-part scientific coding problems with verified Python solutions and discriminative unit tests. It is a full-time 8-week contractor assignment requiring 40 hours a week and four hours of daily overlap with PST, open to residents of a fixed list of countries.

See briefView
New
TuringSTEM AND CODING

Javascript Engineer

Turing hires full-stack JavaScript engineers to build React/Node applications and, in parallel, author the coding tasks and RL training data that frontier labs use to train and benchmark models. The work splits between shipping real features end to end and designing evaluation tasks whose difficulty, tests, and grading hold up under scrutiny.

See briefView
New
TuringOTHER

Music and Audio experts

Turing is hiring trained musicians and audio engineers to listen to short clips and label instrument family and specific instrument for an AI audio dataset. The work runs across a compressed two-day window, fully remote, with observed pay of roughly $150–$160 for a full push across both days.

Rate as listed$150–160 · period unconfirmedView
New
TuringSCIENCES

Web Research Task Author

You build multi-clue web research puzzles that test whether frontier models can find a verifiable answer across the open web, with every clue backed by an exact quote, URL, and verification date. Each task takes roughly four hours and pays around $30 on approval, so read the economics carefully before applying.

Rate as listed$30/taskView
New
TuringOTHER

Music and Audio experts

Turing is hiring trained musicians and audio engineers to label short audio clips by instrument family and specific instrument for a two-day annotation push. Work is fully remote, paid in USD, and gated behind a ~15-minute ear-training qualifying test.

Rate as listed$150–160 · period unconfirmedView
New
TuringSTEM AND CODING

Sr. Data Analyst

A senior data analyst role on a commercial real estate Data Stack project, working inside a Snowflake Medallion architecture to validate, standardize and interpret property, availability, tenancy and comp data. Hired through Turing for a client engagement based in Hyderabad or Gurugram, with 6+ years of analyst experience expected.

See briefView
New
TuringSTEM AND CODING

Senior SWE - Python / TypeScript

Turing is hiring senior engineers to design rubrics and review preference data for frontier-lab training pipelines, while building the Python or TypeScript infrastructure those workflows run on. It's a three-month independent contractor engagement, fully remote within North America, with 40 hours/week preferred and at least six hours of daily overlap with Pacific Time.

See briefView
New
TuringSTEM AND CODING

Engineering Expert

Turing is hiring senior engineers to author simulation-based design problems that current AI agents cannot solve, then build the autograders that score them objectively. The work spans electrical, mechanical, aerospace, control, systems and robotics domains, using open-source simulators like ngspice, OpenFOAM, CalculiX or python-control driven from Python.

See briefView
New
TuringCREATIVE

Video Annotator

Turing is staffing a shot boundary detection project: annotators watch video clips frame-by-frame, mark where each shot starts and ends, label the transition type, and group shots into scene buckets with genre categories. It suits people with editing or production experience who can hold a labeling standard steady across hundreds of clips.

See briefView
New
TuringGENERALIST

Business Analyst

Turing hires business and research analysts to write realistic analytical problems and grade model answers on reasoning, data interpretation, and numerical accuracy. Work is remote, project-based, and paid per task or hour — Turing does not publish a band for this role.

See briefView
New
TuringSTEM AND CODING

SWE – Python Docker

Turing is building verifiable software engineering tasks from real GitHub repository histories to train and evaluate LLMs on bug-fixing work. You pick trending Python repos, Dockerize them into reproducible environments, triage issues, judge test coverage, and confirm a task actually fails and passes as intended.

Rate as listed$100/taskView
New
TuringGENERALIST

Product Mangement SME

Turing is hiring product managers to turn prioritization and roadmap frameworks into deterministic spreadsheet logic and structured parameters that AI tools can follow. You'll build reference RICE/MoSCoW scoring models, PRDs and triage boards, then score and sign off on AI-generated product artifacts.

See briefView
New
TuringSTEM AND CODING

Software Engineering / IT Ops SME

Turing is hiring IT operations and software engineering practitioners to build reference spreadsheet models — SLA trackers, triage escalation logic, asset lifecycle rollups — that AI systems are then graded against. The engagement is a short contractor stint (stated as up to two weeks) at 40 hours/week with 8 hours of PST overlap and weekend on-call availability; pay is not disclosed in the listing.

See briefView
New
TuringGENERALIST

Program Management SME

Turing is hiring experienced program and project managers to build reference assets — Gantt models, RAID logs, dependency formulas, status report templates — that AI systems are trained and scored against. The engagement is short (up to two weeks), contractor-based, 40 hours/week with 8 hours of PST overlap and weekend on-call availability.

See briefView
New
TuringLAW

LLM ServiceNow Product Owner

Turing is contracting ServiceNow product owners to turn real CSM, ITSM and HRSD workflows into scenario-driven training and evaluation tasks for frontier AI models. The work is task authoring in JSON with SQL validation — not platform configuration or app development.

See briefView
New
TuringSTEM AND CODING

Software / Platform / Data Engineer — Agent Trace Collection

Build the end-to-end pipeline that captures GitHub Copilot CLI agent traces from developer machines, scrubs secrets and PII in Azure, and serves a sanitized, queryable dataset to reviewers. It's a three-month contractor assignment at $100/hour, remote within North America, with at least four hours of daily overlap with PST.

Rate as listed$100/hrView
New
TuringSCIENCES

Domain Expert - Statistician / Data Analyst – (jamovi / JASP)

Turing is hiring statisticians and quantitative researchers to build benchmark tasks that test AI models on real statistical work in jamovi or JASP — hypothesis testing, regression, ANOVA, assumption checks, and APA reporting. You construct messy research datasets, run the reference analysis yourself, and write numerical verification criteria precise enough that a grader can mark a model's output right or wrong.

See briefView
New
TuringGENERALIST

Domain Expert - Education / LMS Administrator (Moodle)

Turing is hiring Moodle administrators to build realistic course environments that AI models are then benchmarked against — configuring courses, cohorts, roles, gradebooks and quizzes, then recording reference traces and verification checks. It's a five-week hourly remote engagement; pay is not disclosed in the listing.

See briefView
New
TuringMEDICINE

Domain Expert - Health/EHR Administrator

Turing is hiring healthcare administrators and clinical informatics specialists to build EHR task environments in OpenEMR — patient registration, encounter charting, e-prescribing, and claim workflows — that AI agents are then measured against. You execute each workflow yourself, record the reference trace, and define the database and state checks that decide whether a model's attempt was actually correct.

See briefView
New
TuringSTEM AND CODING

Domain Expert: Mechanical Engineering (FreeCAD)

Turing needs mechanical designers who can execute multi-step parametric modeling tasks in FreeCAD on remote Linux desktops and record clean demonstration trajectories for AI training. You also stress-test automated CAD graders by solving the same brief along alternative valid modeling paths.

See briefView
New
TuringOTHER

Domain Expert - Enterprise Resource Planning (ERP)

Turing is hiring ERP and enterprise-administration practitioners to build realistic task environments — chart of accounts, purchase orders, patient intake forms, contract workflows — that AI agents are benchmarked against. You produce the reference solution, then explain to engineers exactly why a given ledger state or workflow outcome is the correct one.

See briefView
New
TuringSTEM AND CODING

Domain Expert: Electronics & EDA (KiCad)

Turing is hiring KiCad practitioners to execute schematic-capture and multi-layer PCB routing tasks inside remote Ubuntu desktops, recording clean human demonstrations that become reference trajectories for AI agent training and grading. The engagement is pay-per-task with a 60-minute average handling time, initially scoped at five weeks with possible extension.

See briefView
New
TuringSTEM AND CODING

Domain Expert: Mechanical Engineering (FreeCAD)

Turing needs mechanical designers who can execute multi-step parametric modeling tasks in FreeCAD on remote Ubuntu desktops, recording clean human demonstration trajectories for AI training. You also stress-test the automated graders by solving the same brief through alternative valid modeling paths.

See briefView
New
TuringSCIENCES

Domain Expert: Research & Publishing (LaTeX / Overleaf)

Turing is recruiting LaTeX and Overleaf practitioners to execute document-preparation tasks and record clean human demonstrations that train and test AI agents on academic publishing workflows. Work is pay-per-task with an average handling time of 60 minutes, on a five-week engagement that may extend.

See briefView
New
TuringOTHER

Personal Account Generalist Raters (US)

Turing is recruiting US-based generalist raters to evaluate Gemini's personalized responses drawn from the rater's own connected Google apps — Gmail, Calendar, Drive, Photos. The work requires granting app integration permissions to your real personal account, so the data you evaluate against is your own history rather than synthetic samples.

See briefView
New
TuringOTHER

AI Quality Analyst - English

Evaluate a personalization feature for Gemini by writing prompts grounded in your own life, then judging whether the model's use of your Gmail, Search, YouTube, and past-chat data is accurate, natural, and actually helpful. The work requires using your real personal Google account with personal data sources enabled, and writing side-by-side rankings with turn-referenced rationales.

See briefView
New
TuringOTHER

Personal Account Generalist Raters (Korean)

Turing is recruiting US-based Korean-speaking generalist raters to evaluate Gemini's personalized responses drawn from their own connected Google apps — Gmail, Photos, Calendar, Drive. The work requires granting app integration permissions to your real personal account, so the gate is as much about consent and account history as it is about rating skill.

See briefView
New
TuringGENERALIST

Business Analyst

A full-time Business / Product Analyst role in Hyderabad, working on data quality for a commercial real estate system of record where inbound records currently sit in a manual review queue. You'll profile coverage gaps, reconcile cross-source conflicts, and write the rules that decide what a human reviews versus what an agentic workflow can auto-correct.

See briefView
New
TuringSCIENCES

Scientific Coding - Material science and Python

Author materials-science coding problems — a main problem plus at least three connected sub-problems — and implement verified Python solutions with unit tests that frontier models are scored against. Work runs 40 hrs/week for 8 weeks on Turing's Central Task Platform, with 4 hours of daily overlap with PST and QC iteration until tasks clear multi-judge Pass@K criteria.

See briefView
New
TuringSCIENCES

Scientific Coding - Biology and Python

Author multi-part biology coding problems with verified Python golden solutions and discriminative unit tests for Turing's SciCode dataset, built with NVIDIA to train and evaluate frontier models. Full-time contractor work for eight weeks, with four hours of daily overlap with PST and quality gated by automated structure checks and multi-LLM Pass@K review.

See briefView
New
TuringSCIENCES

Scientific Coding - Chemistry and Python

Turing's SciCode project pays chemistry-trained Python programmers to author multi-step scientific coding problems with verified golden solutions and discriminative unit tests, used to train and evaluate frontier models. It is a full-time 8-week contractor engagement with 4 hours of daily PST overlap, open to applicants in a specific list of countries.

See briefView
New
TuringSCIENCES

Scientific Coding - Mathematics and Python

Turing's SciCode project needs mathematicians who can author multi-step scientific coding problems in Python, write the verified golden solution, and build test cases that reliably separate correct model output from plausible-but-wrong output. It is a full-time 8-week contractor engagement with 4 hours of daily PST overlap, open to applicants in a fixed list of countries.

See briefView
New
TuringSCIENCES

Scientific Coding - Physics and Python

Turing's SciCode project needs physicists with a Master's or PhD to author multi-part scientific coding problems, implement verified Python solutions, and design test cases that separate correct model output from plausible-looking errors. It is a full-time 8-week contractor engagement with four hours of daily PST overlap, open to candidates in a specific list of countries.

See briefView
New
TuringSCIENCES

AI Systems Engineer - STEM Workflow

Turing is staffing a senior engineer to build multi-step agentic pipelines for an exploratory mathematics research project, working directly with research mathematicians. It is a four-week contractor assignment at 40 hrs/week with at least four hours of daily overlap with PST.

See briefView
New
TuringGENERALIST

Finance Experts

Turing is hiring finance practitioners to review — not write — finance question-and-answer tasks used as training data for frontier AI labs. You judge whether each item is factually correct, logically sound, clearly worded, and consistent with how the discipline is actually practised.

See briefView
New
TuringSCIENCES

Health/Biology Expert

Turing is contracting biology and health science experts to write graduate-level, text-only biology problems with fully worked solutions and to audit question-answering tasks produced by peers and by AI models. The engagement runs four weeks as a contractor, remote, with at least four hours of daily overlap with PST and weekend on-call availability.

See briefView
New
TuringSCIENCES

Humanities Expert - AI Training Data Reviewer

Turing is contracting humanities-trained reviewers to evaluate AI-generated questions and answers across history, philosophy, literature, linguistics, and the social sciences. The engagement is full-time for four weeks, remote, with four hours of daily overlap with PST.

See briefView
New
TuringSCIENCES

Domain Experts - Scientific Computing Specialist

Turing is hiring a scientific computing specialist to build reusable task templates for benchmarking AI models on GNU Octave and GeoGebra, plus mapped coverage of R, RStudio, SPSS, JMP, JASP, and Scilab/Xcos. You map what each application can actually do, produce realistic base assets (.m scripts, .ggb worksheets, .R files, numerical CSVs, target matrices), and complete reference solutions that define what correct looks like.

See briefView
New
TuringSCIENCES

Domain Experts - Physicist (Modelling & Simulation) Specialist

Turing is hiring computational physicists to build the reusable templates and base assets behind AI benchmark environments for desktop simulation software — Gaussian, OpenFOAM, SU2, Elmer FEM and WRF. You map what each application can actually do, author realistic input decks and datasets, and complete reference solutions that define what correct looks like for engineers who write the scoring code.

See briefView
New
TuringGENERALIST

Business Management Consultant

Turing is contracting consultants to review, rewrite, and score enterprise PowerPoint decks and the prompts that generate them, producing training data for frontier models. The work draws on management-consulting deck craft: structured storylines, MECE logic, executive-ready formatting, and defensible business judgment.

See briefView
New
TuringMEDICINE

Medical Expert – Dermatology

Turing is hiring India-based dermatologists to review dermatology images in a structured annotation tool and to evaluate AI-generated dermatologic text for accuracy, safety, and completeness. The engagement is a 12-week freelance contract at roughly 40 hours per week, with four hours of daily overlap with US Pacific time.

See briefView
New
TuringSTEM AND CODING

Agentic Coding Annotator - Online / Offline Tasks

Turing is hiring contractors under an "Agentic Coding Annotator" title whose posted requirements are those of a production DevOps engineer: GPU infrastructure on GCP, LLM serving with vLLM or Triton, and data pipelines. Expect a two-month contractor engagement at 8 hours per day with a 4-hour overlap with PST, remote, pay undisclosed.

See briefView
New
TuringCREATIVE

Sports & Fantasy Sports SME

Turing is hiring a fantasy sports and sports analytics expert to define the quality bar for AI-generated draft boards, scoreboards, and projection tools. The work is short — up to two weeks, full-time, with weekend on-call — and centres on turning league rules into deterministic, checkable logic.

See briefView
New
TuringSCIENCES

Lifestyle Experts

Turing is hiring lifestyle practitioners to write domain "recipes" that teach a Google Sheets AI assistant how to turn plain-text requests into working interactive apps for scenarios like fitness tracking, event planning, road trips, or pet care. You own each recipe end-to-end — domain rules, discovery questions, a reference Google Sheet, and iteration on the generated app — and every deliverable is checked for authentic, human-written voice.

See briefView
New
TuringMEDICINE

Health & Fitness SME

Turing is hiring strength and conditioning practitioners to encode training logic — periodization, progressive overload caps, e1RM math, volume rollups — into the guardrails and golden assets that fitness-focused AI tools are graded against. It's a full-time contractor engagement of up to 10 weeks, remote, with 8 hours of daily overlap with PST and weekend on-call availability.

See briefView
New
TuringOTHER

Travel & Events SME

Turing is hiring travel and event operations professionals to ground LLM outputs in real itinerary, vendor, and cost-splitting logic. The work is authoring guardrails, validating context files, and signing off on AI-generated trip planners and event trackers as a domain expert.

See briefView
New
TuringGENERALIST

Personal Finance SME

Turing is hiring consumer-finance practitioners to ground LLM behaviour in household money management — budgeting methods, debt payoff sequencing, sinking funds, and net-worth tracking — and to convert that logic into deterministic spreadsheet formulas. The engagement runs up to 10 weeks at 40 hours a week with 8 hours of daily overlap with PST, plus weekend on-call availability.

See briefView
New
TuringLAW

Food & Nutrition SME

Turing is contracting nutrition and culinary operations experts to ground LLM-generated meal planning tools in real dietetic and menu-engineering math. The work is full-time contract for up to 10 weeks, remote, with 8 hours of daily overlap with PST and weekend on-call availability.

See briefView
New
TuringLAW

Home / Interior / Real Estate SME

Turing is hiring real-estate investment, property management, and renovation specialists to ground LLM-built financial tools in correct domain logic. You author guardrails (NOI bounds, vacancy caps, reserve allocations), verify cash-flow and cap-rate formulas, and sign off on AI-generated analyzers and project tools.

See briefView
New
TuringGENERALIST

Operations / Supply Chain SME

Turing is contracting logistics and fleet-operations practitioners to ground LLM-built supply chain tools — route planners, work order managers, vendor trackers — in real dispatch logic. You author the constraints and guardrails, validate context files like dispatch logs and driver rosters, and sign off on whether AI output would survive contact with an actual dispatch desk.

See briefView
New
TuringGENERALIST

HR & People Ops SME

Turing is contracting HR and People Ops practitioners to ground large language models in real HR logic — requisition pipelines, headcount plans, comp bands, and PTO accrual rules. The work is building and reviewing spreadsheet-based 'golden assets' where every formula, KPI definition, and validation rule has to survive audit by another People Ops professional.

See briefView
New
TuringGENERALIST

Sales / Marketing / CRM SME

Turing is contracting domain experts in sports analytics and fantasy-sports mechanics to ground and evaluate LLM outputs on draft boards, scoring engines, and projection tools. The engagement is full-time contractor work for up to 10 weeks, fully remote, with 8 hours of daily overlap with PST and weekend on-call availability.

See briefView
New
TuringCREATIVE

Audio Transcribers - Audio Annotation & Diarization (Arabic)

Turing is building an evaluation-grade dataset of multi-channel, multi-speaker Arabic conversations, and this role is the final QA pass on both the audio and the human-corrected transcripts. You audit channel isolation and recording fidelity, verify verbatim transcription and word-level timestamps in JSON, and issue pass/fail decisions with written justification.

See briefView
New
TuringSTEM AND CODING

Senior Software Engineer – Python (LLM Evaluation & Repository Validation)

Build coding-agent benchmark tasks from production-like Python repositories, then write the visible and hidden test suites that catch an agent completing the work by unsafe means — disabling tests, leaking secrets, or broadening permissions. Turing lists this as a remote contractor assignment of 4–8 weeks, 20+ hours per week with four hours of PST overlap.

See briefView
New
TuringCREATIVE

Domain Experts - Audio Engineer (REAPER /Ardour)

Turing is hiring an audio engineering SME to build the reusable project templates that AI agent-evaluation environments run on — multi-track sessions, REAPER and Ardour project files, effect chains, and mastering references. You define what a technically correct mix, stem, or export looks like and document the reasoning; engineers write the scoring code.

See briefView
New
TuringSTEM AND CODING

Senior Infrastructure / Backend Engineer — GCP Production Systems

Turing is staffing a three-month contract to run production GCP infrastructure behind an AI data pipeline — Cloud Run runners, a final-packaging system, and a QC service customized for Perplexity. You own deployments, incidents, and trainer-facing technical support independently, at full-time hours with 6–8 hours of PST overlap.

See briefView
New
TuringOTHER

Domain Experts - Research & Technical Writing (LaTeX)

Turing is hiring LaTeX-fluent technical writers to build the document environments used to benchmark AI models on real desktop engineering software. You produce multi-file LaTeX repositories, BibTeX databases, custom .cls/.sty packages and TikZ diagrams, then solve reference tasks on record so there is proof of what correct looks like.

See briefView
New
TuringGENERALIST

Mathematics Research Specialist – Algebraic Geometry

Turing is contracting mathematicians with algebraic geometry expertise for a short research engagement focused on proof development and formalizing recurring proof structures into reusable specifications. Work is remote and contractor-based, up to four weeks, with at least four hours of daily overlap with PST.

See briefView
New
TuringCREATIVE

Domain Experts - Creative & Media (Graphic / Media Designer + Video & Audio Editors)

Turing is hiring graphic/media designers and video & audio editors to produce creative assets and evaluate model-generated visual and audio work against structured rating rubrics. The assignment is a 5-week remote contractor engagement at 40 hours per week with at least 4 hours of daily PST overlap.

See briefView
New
TuringSTEM AND CODING

Domain Experts - Aerospace / Flight-Dynamics Engineer

Turing is hiring aerospace and flight-dynamics engineers to author simulation tasks in OpenVSP, XFLR5, FlightGear+JSBSim, OpenRocket and QGroundControl, build reference ('golden') geometries and mission plans, and score model-produced geometry, route and airframe files. It is a 40-hour-per-week contractor assignment, initially five weeks, remote with at least four hours of PST overlap.

See briefView
New
TuringCREATIVE

Photo Editing Specialist

Turing is contracting photo editors and digital artists to produce and retouch images against detailed creative briefs, working across styles from oil and watercolour to pop art and paper craft. Work is remote for roughly 12 weeks, up to 40 hours a week, at an observed rate of around $7/hour, and begins only after your portfolio is approved.

Rate as listed$7/hrView
New
TuringCREATIVE

Senior Artist (Graphic Designer)

Turing is contracting illustrators to produce multi-style digital art — cartoon, manga/anime, pop art, sticker, and sketch — as reference and training assets for foundational model work. The engagement runs up to 12 weeks remote at an observed rate of around $7/hour, with placement contingent on portfolio approval.

Rate as listed$7/hrView
New
TuringGENERALIST

Senior Accounting & Finance Advisors

Turing is recruiting senior finance practitioners — controllers, CFO advisors, finance managers — to answer and evaluate complex accounting questions for an AI assistant workflow built around QuickBooks Online and chat support. The engagement runs up to 40 hours per week for 10 weeks, fully remote, with extension tied to performance.

See briefView
New
TuringCREATIVE

Digital Imaging Specialist/ Digital Artist

Turing is contracting 3D and digital imaging artists to build stylized assets — low-poly, voxel, handcrafted, and material-driven — from written briefs and reference boards, with the output feeding AI training and evaluation work. It is a 12-week remote engagement at an observed rate of around $7/hour, paid in USD, with entry after a portfolio review.

Rate as listed$7/hrView
New
TuringGENERALIST

US Tax Forms Experts

Annotate US tax forms (W-2, 1040, 1099 series, corporate schedules) in Label Studio by drawing bounding boxes around fields and tagging the relationships between them. Turing runs this as a freelance, fully remote engagement of roughly one month at up to 40 hours per week, with extension tied to project needs.

See briefView
New
TuringCREATIVE

Illustrator/Sketcher/Cartoonist

Turing is contracting illustrators and cartoonists to produce digital character artwork from reference images, generating creative variations that preserve character identity across styles. Observed pay is around $5/hour on approved tasks, with a 10–12 week remote engagement at 20–40 hours per week.

Rate as listed$5/hrView
New
TuringSCIENCES

Web Research Specialist- US only

Turing is building an evaluation benchmark for frontier AI browsing agents, and this role writes the problems those agents fail. You start from a verifiable fact, reverse-engineer a research question that is brutally hard to locate, and document a complete, auditable evidence trail including the obvious searches that came up empty.

Rate as listed$60/taskView
New
TuringSCIENCES

Chemistry Expert

Turing contracts credentialed chemists to write high-difficulty tasks and gold-standard solutions used to train and evaluate frontier models, then calibrate other experts' work against a shared rubric. Engagement is remote contractor work, up to 8 weeks, with at least 4 hours a day and 4 hours of overlap with PST.

See briefView
New
TuringSTEM AND CODING

QC and Code Expert

Turing is hiring practicing lawyers to write high-difficulty legal tasks with gold-standard answers, then review and calibrate other experts' work against the rubric. Despite the "QC and Code" title, the stated requirements are legal: 5–15 years of practice, a JD or equivalent, and active bar admission.

See briefView
New
TuringSTEM AND CODING

Sr. Python Engineer

Turing is contracting senior Python engineers to build and maintain the backend APIs that frontier AI labs use for training, evaluation, and benchmarking runs. It is a three-month remote contractor engagement, minimum 20 hours per week with four hours overlapping PST, open to candidates in a specific list of countries.

See briefView
New
TuringLANGUAGES

Audio/Voice/Annotation Trainer - Korean Language

Turing is recruiting native-level Korean voice actors to record scripted lines for AI training and evaluation datasets, capturing specific emotions, personalities and pacing on direction. Work is freelance and remote, requiring a quiet room and professional-grade recording equipment, with delivery-reviewed audio files submitted to spec.

See briefView
New
TuringCREATIVE

QA Specialist - Audio Annotation & Diarization (Portugese)

Turing is staffing a final-review QA layer on an evaluation-grade dataset of multi-channel, multi-speaker conversational audio, where you check both recording fidelity and the human-corrected transcription, diarization, and metadata beneath it. The posting is titled Portuguese but its stated language requirement reads "native proficiency in Spanish (LATAM)" — confirm which language the pod actually needs before you invest time in the screen.

See briefView
New
TuringLANGUAGES

Audio/Voice/Annotation Trainer - French Language

Turing is recruiting French-speaking voice actors to record scripted and directed audio for AI training datasets, working from a professional home studio or in-studio sessions. Work is freelance and project-based, with a delivery review replacing the usual interview round.

See briefView
New
TuringLANGUAGES

Audio/Voice/Annotation Trainer - Italian Language

Turing is recruiting native-level Italian voice actors to record scripted lines — characters, narration, emotional range — for AI speech and multimodal training data. Work is remote and freelance, with your own professional-grade home studio, and starts after a delivery review of your recorded output.

See briefView
New
TuringLANGUAGES

Audio/Voice/Annotation Trainer - Portuguese (Brazil) Language

Record Brazilian Portuguese voice-over from your own studio for AI training and evaluation datasets, performing scripted lines with controlled emotion, pacing and clarity. Turing shortlists on prior voice acting experience and then runs a delivery review of your actual recorded output before onboarding.

See briefView
New
TuringSCIENCES

Web Research Specialist

Turing is building an evaluation benchmark for frontier AI browsing agents, and this role designs research questions those agents cannot solve even with full web access and repeated attempts. You work backwards from a verifiable fact to construct a hard-to-locate question, then document an auditable evidence trail proving the obvious searches fail.

Rate as listed$30/taskView
New
TuringLANGUAGES

Audio/Voice/Annotation Trainer - German Language

Turing is recruiting German-speaking voice actors to record scripted and character-driven audio for AI training datasets, working from a home studio or in-studio sessions. Work is freelance and project-based, with a delivery review replacing the usual interview round.

See briefView
New
TuringLANGUAGES

Audio/Voice/Annotation Trainer - Spanish (Spain) Language

Record Spanish (Spain) voice performances and annotate audio for AI training datasets, working from your own home studio or an in-studio session. Turing shortlists on voice acting credits and equipment, then runs a delivery review on a sample recording before you start.

See briefView
New
TuringGENERALIST

CUA Data Annotator

Turing is hiring CUA (Computer-Using Agent) Data Annotation Trainers to onboard annotators, write SOPs, and audit annotated data for a two-week, full-time contractor pod. The work is equal parts mentoring, documentation, and quality review of computer-agent trajectory data, with four hours of daily overlap with US Pacific time.

See briefView
New
TuringGENERALIST

CUA Quality Analysts

Turing is hiring quality analysts for Computer-Using Agent (CUA) data work — reviewing annotated trajectories, training annotators, and writing the guidelines everyone else works from. It is a full-time, two-week contractor assignment requiring 40 hours per week with four daily hours overlapping Pacific time.

See briefView
New
TuringCREATIVE

Vision Annotators

Turing contracts freelancers to evaluate and annotate AI-generated content — judging accuracy, logic and consistency, then writing the explanation of why a response is right or wrong. Note that while the listing is posted as "Vision Annotators," the stated duties are general LLM evaluation and reasoning work, with a 20-hour weekly minimum and four hours of daily overlap with PST.

See briefView
New
TuringGENERALIST

Business Analyst - Data Annotator (US)

Turing is recruiting US-based writers and analysts to label, classify, and validate the datasets used to train and evaluate frontier AI models. The work is remote contractor engagement at roughly 40 hours a week, with four daily hours overlapping Pacific time and pay not disclosed at listing.

See briefView
New
TuringGENERALIST

CUA Data Annotation Trainer

Turing is hiring contractors to train and quality-check annotators producing computer-use agent (CUA) data — desktop and browser interaction trajectories used to teach models to operate software. You review annotated trajectories, write SOPs and guidelines, run onboarding and calibration sessions, and escalate edge cases with QA and project managers on a 4-week, 40-hour-per-week engagement.

See briefView
New
TuringSTEM AND CODING

Subject Matter Expert : Engineering

Turing is recruiting game systems designers to specify and validate the mechanics behind AI-generated interactive content — progression curves, XP and reward economies, and quest/task board logic. Work is project-based across a 16-week engagement, with each task involving multi-chatbot comparison and a final written assessment.

See briefView
New
TuringOTHER

Music and Audio experts

Turing is hiring trained musicians and audio engineers to listen to short clips and label instrument family and specific instrument for an AI audio dataset. It's a two-day project window, fully remote, with an observed band of roughly $150–160 for a full push across both days and possible extension for strong annotators.

Rate as listed$150–160 · period unconfirmedView
New
TuringSTEM AND CODING

SWE – Python Docker

Turing is building verifiable software engineering tasks from real GitHub repository histories, and needs Python engineers who can Dockerize repos, triage issues, and judge test quality. Pay is observed at $100 per accepted task, on a three-month contractor assignment open to residents of eight listed countries.

Rate as listed$100/taskView
New
TuringSTEM AND CODING

Senior Software Engineer – C++(LLM Evaluation & Repository Validation)

Turing is hiring senior C++ engineers to turn real GitHub issue histories into verifiable software-engineering tasks that test how well LLMs fix bugs in production codebases. The work is hands-on: Dockerize repositories, reproduce issues locally, judge test coverage quality, and flag the cases models still fail.

See briefView
New
TuringFINANCE

Senior Management Consultants (LatAm)

Turing is hiring senior strategy consultants across Latin America to write hard business-strategy tasks and gold-standard answers that train frontier models to reason like an Engagement Manager. You'll also review other experts' work against rubrics and flag ambiguities to the Functional SME Lead, on a contract of up to nine weeks.

See briefView
New
TuringSTEM AND CODING

Senior Backend Engineer (Python, SQL & AI Integration)

Turing is screening for a QA engineer to own end-to-end quality for a billing preparation pipeline that combines a rules engine, AI model outputs, REST APIs, and a human review workbench. The work is Python test automation plus SQL data validation, business-led UAT facilitation, and acting as the deployment gate — based in Hyderabad, full-time.

See briefView
New
TuringFINANCE

Senior Accountants & Auditors (APAC/MEA & NA)

Turing is contracting senior accountants and auditors to write high-difficulty accounting and audit tasks with gold-standard solutions, then review and calibrate other experts' work. The engagement runs up to 9 weeks as a contractor, requiring at least 4 hours per day with 4 hours of overlap with PST.

See briefView
New
TuringSTEM AND CODING

Senior Backend Engineer (Python, SQL & AI Integration)

Build the GCP-native billing pipeline that replaces a manual Excel process for the UC-003 Billing Prep use case, spanning a configurable rules engine, Cloud Run REST services, and PostgreSQL state on Cloud SQL. This is a full-time delivery role based in the Hyderabad office, reporting to a Solution Architect, with bidirectional integration into the Impact billing system and PeopleSoft-sourced BIL files.

See briefView
New
TuringSCIENCES

Subject Matter Expert (SME) – Context Elicitation & Conversation Evaluation

Turing is recruiting clinicians, qualitative researchers, and advisory professionals to review transcripts of conversations between people and AI models, judging whether the model gathered enough context before giving advice. The work is a four-week, full-time contractor assignment with pay-per-task compensation and at least four hours of daily PST overlap.

See briefView
New
TuringSTEM AND CODING

Principal GenAI Engineer

Build and ship the AI layer of a billing preparation pipeline — reimbursable cost matching, anomaly detection, and Gemini-powered copilot features on Google Vertex AI. This is a full-time engineering seat in Hyderabad reporting to a Solution Architect, not a part-time evaluation gig; pay was not disclosed on the listing.

See briefView
New
TuringSTEM AND CODING

Fullstack Developer – AI Applications (GCP)

Turing is contracting fullstack developers to build production web applications on top of LLM inference infrastructure, using Python, Node.js/TypeScript, React/Next.js and GCP. It is a two-month remote contract at 8 hours per day with a required four-hour overlap with PST.

See briefView
New
TuringFINANCE

Senior Financial & Investment Analysts (Europe & LatAm & NA)

Turing is hiring experienced finance professionals to write hard financial-analysis tasks and gold-standard solutions used to train and evaluate frontier AI models. The work is contract, remote, and runs up to nine weeks with a minimum of four hours per day and four hours of daily overlap with PST.

See briefView
New
TuringGENERALIST

Finance, Accounting & Advisory SME Lead

Turing is hiring a senior finance and advisory leader to write and grade high-difficulty expert tasks across markets, corporate finance, accounting/audit/tax, and consulting reasoning for frontier model training. You set the quality bar for a pod of finance SMEs, approve their gold-standard solutions, and escalate rubric ambiguities to the Functional SME Lead.

See briefView
New
TuringSTEM AND CODING

Agentic Coding Annotator - Online / Offline Tasks

Turing is staffing a contractor role that operates GPU infrastructure and production LLM serving stacks on GCP — GKE, vLLM/Triton, Docker, IaC, and data pipelines — in support of frontier-model coding and agentic work. The posting is listed under an "Agentic Coding Annotator" title, but the stated requirements are squarely those of a hands-on DevOps/ML infrastructure engineer, so expect the screen to probe infrastructure depth rather than annotation experience.

See briefView
New
TuringSTEM AND CODING

Agentic Coding Annotator - Online / Offline Tasks

Turing needs experienced software engineers to run realistic coding tasks inside an agentic harness, then rank, grade, and justify model trajectories with evidence. Offline assignments add task authoring, rubric design, and environment calibration on a five-week contract.

See briefView
New
TuringLAW

Lawyers — Senior Expert

Turing is contracting senior lawyers to write high-difficulty legal tasks and gold-standard answers used to train and evaluate frontier AI models, and to calibrate the work of other legal contributors. The engagement runs up to 8 weeks as a contractor, 4–40 hours per week with four hours overlapping PST.

See briefView
New
TuringLANGUAGES

Audio/Voice/Annotation Trainer - English (US) Language

Turing is contracting US English voice actors to record scripted speech for AI training and evaluation datasets, working from a home studio or in-studio sessions. Work is freelance and project-based: interpret scripts, take direction, revise takes, and deliver clean audio files to spec.

See briefView
New
TuringOTHER

Personal Account Generalist Raters (US)

Turing is recruiting US-based generalist raters to evaluate Gemini's personalization quality using their own connected Google accounts — Gmail, Calendar, Drive, Photos and similar. You rate whether model responses retrieve the right personal context, avoid unsupported assumptions, and actually help, then write up why.

See briefView
New
TuringFINANCE

Senior Management Consultants (LatAm)

Turing is hiring senior strategy consultants across Latin America to write hard business-problem tasks and gold-standard solutions that teach AI models to reason like an Engagement Manager or Principal. The work is contract, remote, up to nine weeks, with at least four hours of daily overlap with PST.

See briefView
New
TuringGENERALIST

Senior Marketing Managers (LATAM, US)

Turing is hiring senior marketing leaders to write and grade complex marketing tasks — strategy, segmentation, positioning, channel mix, measurement, budget allocation — used to train and evaluate frontier AI models. Work is contractor-based for up to nine weeks, 4–40 hours per week with at least four hours overlapping PST, and pay is not disclosed on the listing.

See briefView
New
TuringCREATIVE

QA Specialist - Audio Annotation & Diarization (Japanese)

Turing is building an evaluation-grade dataset of multi-channel Japanese group conversations, and this role is the final quality gate on both the audio and the human-corrected transcripts. You review recordings for channel bleed and noise, verify verbatim transcription, word-level timestamps and speaker labels in JSON, and pass or fail each session with written justification.

See briefView
New
TuringSTEM AND CODING

Forward Deployed Engineer — Data Engineering & GenAI

A full-time, hybrid Forward Deployed Engineer role in the Washington, DC metro area, building GenAI applications and data pipelines directly alongside customers. Expect to move between discovery conversations, ETL work, agent and RAG development, and production deployment — often within the same week, and with U.S. citizenship required.

See briefView
New
TuringSTEM AND CODING

Mechanical Engineering

Turing is recruiting power electronics engineers to author and grade advanced PCB design tasks used to benchmark AI design systems. You will build schematics, layouts and manufacturing outputs, then define the verification rubrics that decide whether a model's design actually meets spec.

See briefView
New
TuringSTEM AND CODING

Principal GenAI Engineer

Turing is staffing a Staff/Principal AI Engineer to design and ship production agentic systems — tool-calling agents, RAG pipelines, guardrails, and evaluation harnesses — for Fortune 500 clients. It is a hands-on full-time role based in Bengaluru with three days in office and a daily overlap into US Eastern morning hours; pay is not disclosed on the listing.

See briefView
New
TuringGENERALIST

Business Analyst - US Educational Raters - Gemini

Turing is staffing US-based raters to evaluate Gemini's responses to education-flavored prompts — research tasks, analytical problems, and written explanations — at an observed $30/hour on a 16-week engagement. The unusual gate is that you must already use the Gemini app: at least 10 prior conversations about learning, research, coursework, or skill development, verified before onboarding.

Rate as listed$30/hrView
New
TuringOTHER

Music and Audio experts

Turing is recruiting formally trained musicians and audio engineers to label short audio clips by instrument family and specific instrument for a frontier-AI audio dataset. The work runs across a compressed two-day window, fully remote, with an observed total of roughly $150–160 for a full push across both days.

Rate as listed$150–160 · period unconfirmedView
New
TuringGENERALIST

Domain Expert- Sports

Turing is hiring sports domain experts to write hard prompts about leagues, rules, athletes, and statistics, then judge model answers for factual accuracy and reasoning quality. It is a full-time 8-week contractor assignment requiring 40 hours a week with at least four hours of overlap with US Pacific time.

See briefView
New
TuringGENERALIST

Domain Expert- Politics

Turing is hiring politics specialists to write adversarial prompts and grade LLM answers across political science, governance, elections, public policy and international relations. The assignment is a full-time contractor engagement — 40 hours a week for eight weeks, with at least four hours of daily overlap with PST.

See briefView
New
TuringSCIENCES

Domain Expert - TV Show & Movies

Turing is contracting film and television specialists to write adversarial prompts and grade LLM answers on movies, TV, streaming, awards and entertainment history. It's a full-time, eight-week contractor assignment requiring four hours of daily PST overlap, open to US residents outside California, Connecticut, Massachusetts, New Jersey and Vermont.

See briefView
New
TuringGENERALIST

Domain Expert- Music

Turing is hiring music specialists to write hard prompts and grade LLM answers across theory, history, genre, instruments, and production. It is a full-time contractor assignment — 40 hours a week for 8 weeks, remote, with at least 4 hours of daily overlap with US Pacific time.

See briefView
New
TuringSCIENCES

Domain Expert - Art

Turing is contracting art specialists to write hard prompts and grade LLM answers across art history, fine and visual arts, architecture, design, and museum practice. It's a full-time 8-week contractor assignment at 40 hours per week with at least four hours of daily overlap with PST.

See briefView
New
TuringSTEM AND CODING

Azure DevOps Engineer(LangGraph)

Turing is placing a mid-senior Azure DevOps engineer to build CI/CD pipelines and cloud infrastructure for LangGraph-based agentic AI applications. The role is remote within India, requires 8+ years of DevOps experience, and expects availability within two weeks.

See briefView
New
TuringLAW

Software Engineer – AI Code Evaluation & Benchmarking (US candidates only)

Turing is contracting US-based software engineers to review AI-generated code for correctness, efficiency and maintainability, and to build the benchmarks and rubrics those judgments run on. It is a one-month remote contractor assignment requiring at least 20 hours a week with four hours overlapping PST, and pay is not disclosed in the listing.

See briefView
New
TuringSTEM AND CODING

Board Game Reasoning Expert (AI Training & Evaluation)

Turing is contracting board game and strategy specialists to build and grade reasoning tasks that test how well frontier LLMs handle rules, probability, and multi-step planning. Work is remote and freelance, 20+ hours a week with four hours overlapping PST, on a two-month contract screened by a take-home assessment.

See briefView
New
TuringGENERALIST

Senior Accounting & Finance Advisors

Turing is recruiting senior finance practitioners — controllers, CFO advisors, finance directors — to answer complex accounting and bookkeeping questions inside a QuickBooks Online support context, and to judge whether AI-generated answers hold up. The engagement runs up to 40 hours a week for an initial 10 weeks, fully remote, with selection via Turing's Vetsmith Challenge and a client profile screen.

See briefView
New
TuringSTEM AND CODING

CAD & Simulation Consultant | Engineering Design and Simulation – Aerospace / Mechanical / Electrical Engineering

Turing is hiring senior aerospace, mechanical, and electrical engineers to author and review simulation-led engineering tasks used to train and evaluate frontier AI models. You build CAD models, run analyses in tools like ANSYS, MATLAB/Simulink, SPICE or OpenFOAM, define objective acceptance criteria, and peer-review other engineers' submissions for technical correctness.

See briefView
New
TuringGENERALIST

Small business owners - English Business Document

Turing is recruiting small business owners and operators to write realistic business prompts, run them through several AI chatbots, and rank the responses on clarity, usefulness, and accuracy. Tasks are document-grounded — you supply English-language business files like invoices, spreadsheets, and marketing PDFs to anchor the prompts.

See briefView
New
TuringOTHER

AI Quality Analyst - Portuguese (Portugal)

Evaluate a Gemini personalization feature in European Portuguese by writing multi-turn prompts drawn from your own life and judging how well the model uses your Gmail, Search, YouTube, and chat history. Work is side-by-side ranking with written rationales, at an observed $15/hour, 4–40 hours per week for a three-month contract.

Rate as listed$15/hrView
New
TuringGENERALIST

Small business owners (AI response evaluation) - Japanese Business Document

Turing is recruiting Japanese-speaking small business owners to write realistic business prompts, run them against multiple AI chatbots, and rank the responses on clarity, usefulness and accuracy. You must have genuine Japanese-language business documents — invoices, quotations, spreadsheets, PDFs — that you can use as task inputs.

See briefView
New
TuringGENERALIST

Small business owners (AI response evaluation) - Korean Business Document

Turing is recruiting Korean-speaking small business owners to write realistic business prompts, run them against several AI chatbots, and judge which responses would actually hold up in day-to-day operations. The work is project-based over roughly 10 weeks and requires you to supply your own Korean-language business documents — invoices, spreadsheets, contracts, PDFs — as task inputs.

See briefView
New
TuringGENERALIST

Small business owners (AI response evaluation) - Spanish Business Documents

Turing needs small business owners and operators who hold real business documents in Spanish to build prompts from those documents, run them through several AI chatbots, and rank the answers. Work is project-based over roughly 16 weeks, fully remote and asynchronous, with pay undisclosed on the listing.

See briefView
New
TuringLANGUAGES

Audio/Voice/Annotation Trainer - Spanish (Latam) Language

Turing is recruiting Latin American Spanish voice actors to record scripted lines — character work, narration, emotional range — for speech and multimodal AI training datasets. Work is freelance and remote, and you supply your own studio-grade recording setup unless you're booked for an in-studio session.

See briefView
New
TuringLANGUAGES

Audio/Voice/Annotation Trainer - French Language

Record French-language voice performances for AI training datasets — scripted lines, character work, and emotional range delivered from your own studio. Turing shortlists on prior voice acting experience and then runs a delivery review on a sample recording before onboarding.

See briefView
New
TuringLANGUAGES

Audio/Voice/Annotation Trainer - Arabic Language

Record Arabic voice performances — character reads, narration, dubbing and ADR-style lines — for datasets used to train and evaluate speech and multimodal AI models. Work is freelance, fully remote from your own treated recording space, with delivery-reviewed batches rather than fixed hours.

See briefView
New
TuringLANGUAGES

Audio/Voice/Annotation Trainer - Japanese Language

Turing is recruiting Japanese-speaking voice actors to record scripted lines — character work, narration, emotional range — for AI speech and multimodal training data. Work is remote and freelance, with delivery from your own studio or an in-studio session, and screening centres on your reel, your recording setup, and your ability to take direction across takes.

See briefView
New
TuringLANGUAGES

Audio/Voice/Annotation Trainer - Korean Language

Turing is recruiting Korean-speaking voice actors to record scripted performances that train and evaluate speech and multimodal AI systems. Work is remote and freelance, recorded from your own treated home studio, with delivery reviewed against audio-quality and performance specs before you start on live batches.

See briefView
New
TuringOTHER

AI Quality Analyst (Personalization) - German

Evaluate a Gemini personalization feature in German by writing multi-turn prompts drawn from your own life, then judging whether the model's use of your Gmail, Search, YouTube and chat history is grounded, naturally integrated, and genuinely helpful. The role requires connecting your real personal Google account — not a test account — and writing defensible side-by-side rationales that cite specific turns.

Rate as listed$15/hrView
New
TuringOTHER

AI Quality Analyst (Personalization) - Dutch

Evaluate a Gemini personalization feature in Dutch by writing prompts drawn from your own life and judging how well the model uses your real Gmail, Search, YouTube and chat history. Work is side-by-side comparison of two responses plus written rationales that cite specific conversation turns.

Rate as listed$20/hrView
New
TuringOTHER

AI Quality Analyst (Personalization) - Russian

Evaluate a Gemini personalization feature in Russian by writing prompts drawn from your own life, then judging whether the model's use of your Gmail, Search, YouTube and chat history is grounded, natural and genuinely helpful. The role requires connecting your primary personal Google account — not a test account — and writing side-by-side rationales that cite specific conversation turns.

Rate as listed$15/hrView
New
TuringOTHER

AI Quality Analyst (Personalization) - Polish

Evaluate a Gemini personalization feature in Polish by writing prompts drawn from your own life and judging whether the model's use of your Gmail, Search, YouTube and chat history is grounded, natural, and genuinely useful. Contract work through Turing at an observed $20/hr, one month, 4–40 hours a week with four hours overlapping PST.

Rate as listed$20/hrView
New
TuringOTHER

AI Quality Analyst (Personalization) - Vietnamese

Evaluate a personalization feature for Gemini in Vietnamese, writing multi-turn prompts drawn from your own Gmail, Search, YouTube and chat history and judging whether the model's use of that data is grounded, natural, and genuinely helpful. Work is side-by-side model comparison with written rationales, at an observed rate of $15/hour on a 3-month contractor engagement.

Rate as listed$15/hrView
New
TuringOTHER

AI Quality Analyst (Personalization) - Thai

Evaluate a Gemini personalization feature in Thai by writing multi-turn prompts drawn from your own life, then judging how well the model grounds and integrates data from your personal Gmail, Search, YouTube and chat history. The role requires using your primary personal Google account — not a test account — and writing defensible side-by-side rationales that cite specific turns.

Rate as listed$15/hrView
New
TuringOTHER

AI Quality Analyst (Personalization) - Hindi

Evaluate a Gemini personalization feature in Hindi by writing prompts drawn from your own life, then judging whether the model's use of your Gmail, Search, YouTube and chat history is grounded, natural and genuinely helpful. The project requires using your real personal Google account with personal data sources enabled, and running side-by-side comparisons with written rationales that cite specific conversation turns.

Rate as listed$15/hrView
New
TuringOTHER

AI Quality Analyst (Personalization) - Turkish

Evaluate a Gemini personalization feature in Turkish by writing prompts drawn from your own life, then judging how well the model grounds, integrates, and uses your Gmail, Search, YouTube, and past chat data. Work is side-by-side model ranking with written Turkish-context rationales that cite specific conversation turns, at an observed $15/hr on a three-month contract.

Rate as listed$15/hrView
New
TuringOTHER

AI Quality Analyst (Personalization) - Indonesian

Evaluate a Gemini personalization feature in Indonesian by writing prompts drawn from your own life and judging how well the model uses your Gmail, Search, YouTube and past chat history. Work is side-by-side ranking with written rationales, 30–40 hours a week with four hours overlapping PST, at an observed rate of $15/hour.

Rate as listed$15/hrView
New
TuringOTHER

AI Quality Analyst (Personalization) - Korean

Evaluate a personalization feature for Gemini in Korean by designing prompts drawn from your own life, then judging how well the model grounds, integrates, and uses your Gmail, Search, YouTube, and chat history. Work is side-by-side model comparison with written Korean-context rationales, and it requires using your real personal Google account rather than a test account.

Rate as listed$15/hrView
New
TuringSTEM AND CODING

AI Quality Analyst (Gemini) - Chinese

Evaluate a personalization feature for Gemini in Chinese by writing multi-turn prompts drawn from your own Gmail, Search, YouTube and chat history, then judging how well the model grounds its claims about you. Work is side-by-side comparison of two responses with written rationales that cite specific turn numbers.

Rate as listed$15/hrView
New
TuringOTHER

AI Quality Analyst (Personalization) - Spanish

Evaluate a Gemini personalization feature in Spanish by writing multi-turn prompts drawn from your own life, then judging whether the model's use of your Gmail, Search, YouTube and chat history is grounded, natural, and genuinely helpful. The role requires connecting your primary personal Google account — not a test account — and writing side-by-side rationales that cite specific conversation turns.

Rate as listed$15/hrView
New
TuringOTHER

AI Quality Analyst (Personalization) - Portuguese

Evaluate a Gemini personalization feature in Portuguese by writing prompts drawn from your own life, then judging how well the model uses your Gmail, Search, YouTube and chat history. Work is side-by-side ranking of two responses with written rationales that cite specific conversation turns.

Rate as listed$15/hrView
New
TuringOTHER

AI Quality Analyst (Personalization) - Arabic

Evaluate a Gemini personalization feature in Arabic by writing prompts drawn from your own life, then judging whether the model's use of your Gmail, Search, YouTube and chat history is grounded and natural. Work is remote contractor, 3 months, at an observed $15/hr, and requires connecting your primary personal Google account rather than a test account.

Rate as listed$15/hrView
New
TuringOTHER

AI Quality Analyst - English

Evaluate a Gemini personalization feature by writing multi-turn prompts drawn from your own life, then judging whether the model's use of your Gmail, Search, YouTube and chat history is grounded, naturally integrated, and genuinely helpful. The work requires using your primary personal Google account with real data sources enabled, and writing defensible side-by-side rationales that cite specific turn numbers.

See briefView
New
TuringSTEM AND CODING

Technical Content Writer

Write technical documentation and analytical write-ups for frontier AI training data work at Turing, using JSON datasets and light Python to support reproducible analysis. Contractor engagement, fully remote from a defined list of countries, with 20+ hours per week and four hours of PST overlap.

See briefView
New
TuringSTEM AND CODING

MLE Bench – Data Analyst

Turing is contracting experienced data analysts to build and stress-test evaluation scenarios for MLE Bench, a benchmark measuring how well AI systems perform real machine-learning engineering work. The work is hands-on Python and SQL analysis of training, inference, and evaluation outputs, with at least 20 hours per week and four hours of daily overlap with PST.

See briefView
New
TuringSTEM AND CODING

MLE Bench – ML Engineers

Turing is contracting experienced ML engineers to build and solve MLE Bench–style evaluation tasks on production-grade machine learning codebases. The work is hands-on Python: training, evaluation and inference pipelines, dataset and metric preparation, and debugging realistic ML systems so frontier labs can measure how well AI agents handle real engineering work.

See briefView
New
TuringSTEM AND CODING

SWE Bench – Data Engineer/Data Scientist

Turing is hiring data engineers and data scientists to build SWE Bench-style evaluation tasks from real data pipelines and data science codebases. You write and validate Python data workflows that frontier AI models are then tested against, and review peers' tasks for correctness and reproducibility.

See briefView
New
TuringLAW

Prompt & Verifier

Turing is hiring contractors to write and stress-test prompts and policies for AI agents that call third-party tools like Slack and PayPal, then evaluate whether the model's tool calls and outputs actually comply. The work sits between prompt engineering, operational policy drafting, and safety review, with a 20-hour weekly minimum and four hours of daily overlap with PST.

See briefView
New
TuringOTHER

Business Analyst (Tagalog Language)

Turing hires bilingual English–Tagalog analysts to write analytical prompts, reference answers, and explanations that train large language models on reasoning tasks. The work is remote contract at 40 hours per week with 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringOTHER

Business Analyst (Bahasa Indonesian Language)

Turing is recruiting bilingual Bahasa Indonesia/English analysts to write, solve and critique analytical and logic problems used to train large language models. The work is full-time contract (up to 12 months), fully remote, and requires 2–5 hours of daily overlap with US Pacific time.

See briefView
New
TuringOTHER

Business Analyst (Thai Language)

Turing is hiring Thai-English bilingual analysts to write, solve, and critique analytical reasoning tasks used to train large language models. The work is remote freelance at 40 hours per week, with 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringOTHER

Business Analyst (Russian Language)

Turing is recruiting Russian-English bilinguals to write and evaluate analytical reasoning tasks — sales-trend breakdowns, constraint puzzles, claim verification — used to fine-tune large language models. It is a full-time contractor engagement, up to 12 months, remote, with 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringOTHER

Business Analyst (Finnish Language)

Turing is recruiting Finnish–English bilinguals to write analytical prompts, solve reasoning puzzles, and correct model answers with written explanations that teach the model why it was wrong. It's a one-month contractor engagement at up to 40 hours a week, fully remote, with 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringOTHER

Business Analyst (Malay Language)

Write and evaluate analytical reasoning tasks in Malay and English to train large language models on data interpretation, logic puzzles, and claim verification. Turing runs this as a full-time freelance contract, remote, with 2–5 hours of daily overlap with US Pacific hours.

See briefView
New
TuringOTHER

Business Analyst (Hebrew Language)

Turing is recruiting Hebrew-English bilinguals to write and critique analytical reasoning tasks — data interpretation, logic puzzles, claim verification — used to fine-tune large language models. The commitment is 40 hours a week, fully remote, with 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringOTHER

Business Analyst (Turkish Language)

Turing is recruiting Turkish–English bilinguals to write and grade analytical reasoning tasks — sales-trend breakdowns, constraint puzzles, claim verification — that are used to train and correct large language models. The work is fully remote and freelance, with a stated preference for ~40 hours a week and 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringOTHER

Business Analyst (Korean Language)

Turing needs Korean–English bilingual analysts to write analytical problems, logic puzzles, and data-interpretation tasks that expose where large language models reason poorly. You produce the correct answer plus a written explanation the model can learn from, working roughly 40 hours a week with 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringOTHER

Business Analyst (Japanese Language)

Turing contracts bilingual Japanese–English analysts to write reasoning problems, data-interpretation tasks and logic puzzles that expose where large language models fail, then supply the correct answer and a defensible explanation. It is full-time freelance, fully remote, with 2–5 hours a day of overlap with US Pacific time.

See briefView
New
TuringOTHER

Business Analyst (Danish Language)

Turing is recruiting Danish–English bilingual analysts to write, solve, and critique analytical reasoning tasks used to train large language models. The work is freelance, fully remote, and expects roughly 40 hours a week with 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringOTHER

Business Analyst (Dutch Language)

Turing contracts Dutch-English bilingual analysts to write, solve, and critique analytical and logical-reasoning problems used to train large language models. The work is remote freelance, up to 40 hours a week, with 2–5 daily hours overlapping US Pacific time; pay is not disclosed in the listing.

See briefView
New
TuringOTHER

Business Analyst (Arabic Language)

Turing contracts bilingual Arabic–English analysts to write, solve, and critique analytical and logic problems that train frontier language models. The work is full-time contractor hours with a daily overlap with US Pacific time, and pay is negotiated per engagement rather than published.

See briefView
New
TuringOTHER

Business Analyst (Italian Language)

Turing is recruiting Italian–English bilingual analysts to write, solve, and critique analytical and logic problems used to train large language models. The work is full-time contract, fully remote, and requires 2–5 hours of daily overlap with US Pacific hours.

See briefView
New
TuringOTHER

Business Analyst (German Language)

Turing contracts German-English bilingual analysts to write and grade analytical reasoning tasks — data interpretation, constraint puzzles, research-verified claims — that are used to train and evaluate frontier language models. It is a full-time (40 hrs/week) remote contractor assignment requiring 2–5 hours of daily overlap with US Pacific hours; pay is not published on the listing.

See briefView
New
TuringOTHER

Business Analyst (Swedish Language)

Turing is recruiting Swedish–English bilinguals to write and validate analytical reasoning tasks — data interpretation, logic puzzles, claim verification — used to train and evaluate large language models. It is a full-time freelance contract, fully remote, with 2–5 hours of daily overlap with US Pacific hours; pay is not disclosed on the listing.

See briefView
New
TuringOTHER

Business Analyst (French Language)

Turing contracts French–English bilingual analysts to write and grade analytical reasoning tasks — data interpretation, logic puzzles, claim verification — used to fine-tune large language models. It is a full-time freelance commitment of 40 hours a week with 2–5 daily hours overlapping US Pacific time; pay is negotiated per contract and not published.

See briefView
New
TuringOTHER

Business Analyst (Chinese Language)

Turing contracts bilingual English–Chinese analysts to write and grade analytical reasoning tasks — sales-trend breakdowns, constraint puzzles, claim verification — used to fine-tune frontier language models. It's a 40-hour-per-week contractor assignment, fully remote, with 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringGENERALIST

Image/Video Annotator

Annotate short video clips and still images for multimodal model training — labelling subjects, actions, scene transitions, spatial relationships, mood and lighting against detailed guidelines. Freelance and fully remote through Turing, with a stated commitment of 4 hours/day, 20 hours/week and 4 hours of overlap with PST; pay is not disclosed in the listing.

See briefView
New
TuringSTEM AND CODING

Senior Software Engineer – LLM Evaluation (US/Canada/WEU based)

Turing contracts experienced software engineers to build and grade the code data that frontier labs train and benchmark on — writing reference solutions, correcting model output, and designing automatic verifiers across Python, JavaScript/React, C/C++, Java, Rust, and Go. Contractor engagement, 10–40 hrs/week, one-month initial term with extensions based on performance, restricted to candidates based in the US, Canada, or Western Europe.

See briefView
New
TuringSTEM AND CODING

Dockerfile Data Validation Engineer

Build Docker-based data-validation pipelines — Dockerfile LABEL metadata, schema and integrity checks in Python/Bash, and fail-on-bad-data gates in CI/CD — as reference work for frontier AI labs. Short contractor assignment (2–4 weeks) with 20–40 hrs/week and 4 hours of daily PST overlap, restricted to a specific list of countries.

See briefView
New
TuringLANGUAGES

Spanish Voice Actors (Studio-Grade Recording Experience)

Turing is contracting experienced Spanish voice actors to record one finished hour of audio from supplied scripts, delivered as clean, mastered files. The posting states USD 200 per finished hour plus a one-time USD 20 onboarding fee, with a 30-second demo recording as the final screening step.

See briefView
New
TuringFINANCE

Finance Expert (US based)

Turing is hiring US-based finance practitioners to stress-test large language models on capital markets, banking, and accounting work, and to write the rubrics researchers use to score model output. Work is remote and asynchronous at 10–30 hours per week, with observed pay around $100/hour and an initial engagement of roughly one month.

Rate as listed$100/hrView
New
TuringMEDICINE

Resident Medical Specialist (MD/DO)

Turing is recruiting licensed physicians in active clinical practice to design evaluation frameworks that test how AI models reason through real medical problems. The engagement is remote and flexible, up to 30 hours per week, initially scoped at one month with extensions based on fit.

See briefView
New
TuringSCIENCES

Engineering Manager

Turing is hiring a delivery leader to run 20+ person teams of developers producing SFT and RLHF training data for frontier AI labs. The work is part engineering management, part hands-on code review, with direct accountability to researcher clients for dataset quality, throughput, and cost.

See briefView
New
TuringSTEM AND CODING

Python Machine Learning Engineer

Turing is contracting experienced Python ML engineers to build end-to-end machine learning solutions — data pipelines, model design, deployment, and monitoring — for frontier AI labs and enterprise clients. It's a three-month remote contractor assignment requiring 20–40 hours per week with at least four hours overlapping US Pacific time.

See briefView
New
TuringSTEM AND CODING

LLM Go Developer

Review and correct Go code produced by an AI system, judging whether it is idiomatic, correct, and safe to ship. Turing runs this as a flexible contract engagement with two internal interviews (a 60-minute technical plus a short cultural and terms conversation).

See briefView
New
TuringSCIENCES

Ph.D. / Postdoctoral / Master’s Expert

Turing contracts graduate-level STEM specialists — physics-heavy — to write original problems that break frontier language models and to document the correct reasoning step by step. Work is fully remote, freelance, and organised around evaluation benchmarks spanning early undergraduate through PhD-level curricula.

See briefView
New
TuringSTEM AND CODING

Senior Software Engineer – C#(LLM Evaluation & Repository Validation)

Turing is hiring senior C# engineers to turn real GitHub issue histories into verifiable software engineering tasks that LLMs are then tested against. Day-to-day work is hands-on: Dockerizing repositories, triaging issues, judging test coverage quality, and reproducing bugs locally to see where models fail.

See briefView
New
TuringSTEM AND CODING

Senior Software Engineer – Go (LLM Evaluation & Repository Validation)

Build verifiable software-engineering tasks from real Go repository histories so LLMs can be trained and scored on genuine bug-fixing work. Day-to-day means triaging GitHub issues, Dockerising repos, judging test coverage quality, and running codebases locally to see where models break.

See briefView
New
TuringSTEM AND CODING

Senior Software Engineer – Ruby (LLM Evaluation & Repository Validation)

Turing is building verifiable software-engineering tasks from real Ruby repository histories, and needs senior engineers to set up environments, triage issues, and judge whether a task actually tests an LLM. Day-to-day is hands-on: Dockerizing repos, reading test suites critically, and running bug-fix scenarios locally to see where models break.

See briefView
New
TuringSTEM AND CODING

Senior Software Engineer – Rust (LLM Evaluation & Repository Validation)

Turing is hiring senior Rust engineers to turn real GitHub issue histories into verifiable software engineering tasks that LLMs can be trained and tested against. Day-to-day work is hands-on: Dockerizing repositories, triaging issues, judging test coverage, and running candidate fixes locally to see where models break.

See briefView
New
TuringSTEM AND CODING

Senior Software Engineer – C++ (LLM Evaluation & Repository Validation)

Turing is building verifiable software-engineering tasks from real C++ repository histories, and needs tech-lead-level engineers to triage issues, containerize builds, and judge whether a task genuinely tests an LLM. The work is hands-on: you clone real projects, get them building in Docker, reproduce bugs, and assess whether the test suite actually proves a fix is correct.

See briefView
New
TuringSTEM AND CODING

Senior Software Engineer – Python (LLM Evaluation & Repository Validation)

Build verifiable software-engineering tasks from real public repository histories so LLMs can be trained and measured against them — triaging GitHub issues, Dockerizing repos, and judging whether a test suite actually proves a fix. Contractor assignment through Turing, fully remote, 20–40 hrs/week with four hours overlapping PST.

See briefView
New
TuringSTEM AND CODING

Senior Software Engineer – LLM Evaluation

Turing contracts experienced software engineers to build and evaluate code datasets that train and benchmark frontier LLMs, working across Python, JavaScript/React, C/C++, Java, Rust, and Go. The work mixes writing reference solutions, critiquing model-generated code, and designing automated verification for software engineering tasks.

See briefView
New
TuringSCIENCES

Chemistry Expert (PhD/Master's)

Turing contracts chemistry postgraduates to write, solve, and annotate advanced chemistry problems — organic mechanisms, equilibrium, thermodynamics, electrochemistry — used to train and evaluate frontier language models. Work is fully remote, freelance, and multimodal: your answers pair written reasoning with structures, equations, and energy diagrams.

See briefView
New
TuringSCIENCES

Biology Expert (PhD/Master’s)

Write original biology problems designed to break large language models, then supply step-by-step reference solutions rigorous enough to serve as ground truth. Work is freelance and fully remote, spanning undergraduate through PhD-level topics across molecular biology, genetics, physiology, and biochemistry.

See briefView
New
TuringSCIENCES

Mathematics Expert (Master’s/Ph.D.)

Turing contracts mathematicians at Master's, Ph.D., or postdoc level to write original problems, grade model-generated proofs, and formalize arguments in Lean for frontier LLM training and evaluation. Work is fully remote and freelance, with a minimum of 20 hours per week and four hours of daily overlap with PST.

See briefView
New
TuringSCIENCES

Physics Expert (PhD/Master's)

Turing contracts physics graduate students, PhDs, and postdocs to write original problems that break frontier language models, then document the full reasoning path to the correct answer. Work is remote, freelance, and paid per project; Turing does not publish a rate for this listing.

See briefView
New
TuringSTEM AND CODING

Senior Software Engineer – LLM Evaluation & Repository Validation

Build verifiable software engineering tasks from real public repository histories so LLMs can be trained and measured against genuine bug-fixing work. Day-to-day means triaging GitHub issues, Dockerizing repos, judging test quality, and running codebases locally to see where models actually fail.

See briefView
New
TuringGENERALIST

Business Analyst (Finance)

Turing is recruiting finance practitioners in India to grade AI model outputs on corporate finance, M&A, investment, and accounting tasks, and to write the rubrics those grades are based on. It's a freelance contract of roughly one month at 10–30 hours per week, fully remote, with pay not disclosed in the listing.

See briefView
New
TuringMEDICINE

Medicine Physician (MD/DO/Doctoral study/PhD)

Turing is recruiting licensed physicians in active clinical practice to build evaluation frameworks and clinical test scenarios for frontier AI models. The work is remote, flexible up to 30 hours per week, with an initial one-month engagement that can extend based on fit and output quality.

See briefView
New
TuringOTHER

Business Analyst (Vietnamese Language)

Turing is hiring Vietnamese-English bilingual analysts to write and grade analytical problems — data interpretation, logic puzzles, claim verification — used to fine-tune large language models. It's a two-month contractor assignment at up to 40 hours per week, remote, with 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringSTEM AND CODING

Senior LLM Engineer

A hands-on, full-time engineering role in India building production Generative AI systems — RAG pipelines, agent workflows, and LLM services in Python with LangChain. Turing screens for 7–12 years of overall software or ML engineering experience with at least a year of real LLM work shipped to production; pay is not disclosed in the posting.

See briefView
New
TuringOTHER

Business Analyst (Turkish Language)

Turing contracts Turkish–English bilingual analysts to write analytical problems, model answers, and detailed reasoning explanations used to train large language models. Work is fully remote and freelance, with up to 40 hours a week expected and 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringOTHER

Business Analyst (Norwegian Language)

Turing is recruiting Norwegian–English bilingual analysts to write, solve and critique analytical reasoning tasks used to fine-tune large language models. Work is fully remote and freelance, with up to 40 hours a week and 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringOTHER

Business Analyst (Japanese Language)

Turing contracts bilingual (Japanese/English) analysts to write analytical problems, model answers, and detailed critiques used to fine-tune large language models. Work is remote and freelance on a two-month contract, with 20–40 hours per week and 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringOTHER

Business Analyst (French Language)

Turing contracts French–English bilingual analysts to write analytical prompts, reasoning puzzles, and corrected explanations that are used to train and evaluate large language models. The work is remote contract labour on a two-month engagement, with 20–40 hours weekly and 2–5 hours of daily overlap with Pacific time.

See briefView
New
TuringOTHER

Business Analyst (Spanish Language)

Turing is contracting bilingual English/Spanish analysts to write analytical prompts, reference answers, and step-by-step explanations that train large language models on reasoning, data interpretation, and logic puzzles. Work is remote and contract-based (2 months, 20/30/40 hrs per week) with 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringSTEM AND CODING

LLM C/ C++ Developer

Review and validate C/C++ code produced by an AI system, correcting errors and judging whether solutions meet real production standards. Contract work through Turing with flexible hours, aimed at engineers who can defend a code review in writing and work in large codebases.

See briefView
New
TuringOTHER

Content Analyst

Turing contracts writers and editors to build, decompose, and fact-check text used to train large language models. Day to day you summarise long source material, break it into logical blocks, verify claims through online research, and construct puzzles and prompts that stress-test model reasoning.

See briefView
New
TuringSCIENCES

Research Analyst - Advanced Math

Turing contracts analysts to write and solve math and logic problems that current language models get wrong, then explain the correct reasoning step by step so the model can learn from it. Work is fully remote and freelance, with 20–40 hours a week and 2–5 daily hours overlapping US Pacific time.

See briefView
New
TuringSTEM AND CODING

Python + Full-Stack (JS) Developer

Turing contracts Python and JavaScript/TypeScript developers to produce reference code, rank model outputs, and build task-specific datasets for frontier LLM training. The work is remote and freelance, with a 20-40 hr/week commitment and a required four-hour daily overlap with PST.

See briefView
New
TuringSTEM AND CODING

LLM Java Developer

Turing places experienced Java engineers on frontier-model training projects: writing reference solutions, ranking model outputs, and building SFT datasets with written rationales. It is a one-month contractor assignment, fully remote, with a 20–40 hr/week commitment and four hours of daily overlap with PST.

See briefView
New
TuringSTEM AND CODING

JavaScript / TypeScript Full-Stack Developer

Turing contracts JavaScript/TypeScript developers to write reference code, rank model outputs, and build supervised fine-tuning datasets for frontier AI labs. Work is fully remote and contract-based, with a minimum of 20 hours per week and a four-hour daily overlap with PST.

See briefView
New
TuringSTEM AND CODING

Data Scientist/Analyst

Turing contracts data scientists and analysts who write Python to build evaluation tasks, rank model responses on data-analysis problems, and produce reference answers with written reasoning. Work is fully remote and freelance, with a minimum of 20 hours per week and four hours of daily overlap with PST.

See briefView
New
TuringSTEM AND CODING

Senior Python Developer

Write and review Python code that becomes training and evaluation data for a frontier LLM lab, and rank competing model outputs on correctness, style, and reasoning. Contract work through Turing: fully remote, 20–40 hours per week with a four-hour PST overlap, initially a one-month engagement.

See briefView
New
TuringOTHER

Personal Account Generalist Raters (US)

Turing is recruiting US-based generalist raters to evaluate Gemini's personalized responses drawn from your own connected Google apps — Gmail, Calendar, Drive, Photos. The work requires granting app integration permissions so the model can retrieve real personal context, then judging whether its answers are accurate, relevant, and appropriately personalized.

See briefView
New
TuringGENERALIST

LLM Annotator - Master's Degree

Turing is hiring master's-qualified annotators to write hard prompts over structured data and judge whether model answers hold up against the source. It's a four-week contractor assignment, remote, 20–40 hours per week with at least four hours overlapping PST.

See briefView