DevOps Engineer - AI Model Evaluator - AI Trainer
$400 per monthMercor Inc
About the Role Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. Contributors help evaluate and improve frontier AI coding models through structured technical assessments. The work focuses on realistic infrastructure engineering workflows and model evaluation. Spots are limited and filling quickly on a first come, first serve basis. What You'll Do Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-generated implementations involving cloud platforms, Kubernetes, CI/CD systems, observability, and infrastructure automation. Identify bugs, edge cases, reliability issues, and failure modes. Compare outputs from multiple frontier models and assess their strengths and weaknesses. Apply professional engineering judgment to realistic infrastructure engineering scenarios. Time Commitment Sprint based project that runs in 12-24 hour stretches based on client requirement. Compensation $400 per accepted task. Typical tasks take approximately 2–3 hours after ramp-up. Compensation is tied to accepted work. Who Should Apply 2+ years of professional DevOps, SRE, or Cloud Engineering experience. Experience with AWS, Azure, GCP, Kubernetes, Terraform, CI/CD pipelines, or observability tooling. Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools. Ability to evaluate model-generated infrastructure and reliability engineering solutions. Experience supporting production-scale systems is preferred. #J-18808-Ljbffr
$85 per hour
...technical talent with leading AI research labs. Headquartered... ...Dorsey . Position: DevOps / SRE / Cloud Engineer (Coding Agent Experience)... ...coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-generated implementations involving...SuggestedContract workSummer workRemote work- Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. You will help evaluate frontier AI coding models by performing infrastructure engineering tasks and reviewing model-generated implementations on cloud platforms, Kubernetes,...Suggested
$50 - $75 per hour
A leading tech company based in Australia is seeking an AI Model Evaluator on a contract basis. The role involves evaluating AI-generated responses, writing prompts, and providing justifications based on specific criteria. Ideal candidates will hold a Master's degree in...SuggestedHourly payContract work$90 per hour
...technical talent with leading AI research labs. Headquartered in... .... Position: Frontend Engineer — Web Replication Preference Rater... ...fidelity in detail, including box model, typography, color, image... ...that merely look correct. Evaluate responsive behavior and semantic...SuggestedContract workSummer workLocal areaRemote work$80 - $120 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...: BI dashboards / performance reporting Evaluator Type: Contract Compensation: $... ...structured written feedback to improve AI model outputs. Collaborate with teams to ensure...SuggestedContract workSummer workWork at officeRemote work$85 per hour
...and technical talent with leading AI research labs. Headquartered in San... ...Jack Dorsey . Position: iOS Engineer (Coding Agent Experience) Type:... ...AI coding agents to complete and evaluate complex engineering tasks. Review model-generated mobile application code...Contract workSummer workRemote work$15 - $20 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...tools . Generate high-quality human evaluation data by identifying response strengths, areas... ...and completeness of responses. Ensure model responses align with expected conversational...Contract workSummer workRemote work$218.5k - $288k
...Scientist specializing in Small Language Models and AI Training, you will lead research and... .... You will work closely with research, engineering, and product teams to advance model training... ...language models.Design, implement, and evaluate model training experiments to improve...Work at officeFlexible hours3 days per week$180k - $225k
As a Software Engineer on the ML Infrastructure team, you will design... ...engineers to integrate and optimize models for production and research... ...to ensure a fair and thorough evaluation of all applicants.About Us:At... ...is to develop reliable AI systems for the world's most important...Full time$144k - $187k
...power MSCI's quantitative risk and factor model research. The team develops the systems... ...research infrastructure that integrates AI to accelerate model development, validation... ...Collaborating cross-functionally with research, engineering, and data teams Presenting complex...Flexible hours$30 per hour
...A leading platform for AI training is seeking advanced Japanese speakers to join as... ...will complete various AI training tasks and evaluate AI responses in Japanese. Compensation is... .... Join a dynamic environment aimed at improving powerful AI models. #J-18808-Ljbffr...Hourly payRemote workWork from homeFlexible hours$213k - $263k
...across 15+ U.S. states. The mission of the Waymo AI Foundations team is to develop machine learning... ..., learning from demonstration, generative modeling, Bayesian inference, hierarchical learning, and robust evaluation. In this hybrid role, you will report to a...Full timeTemporary workRemote work$150k - $200k
...This is a founding-team opportunity to join the macOS DevOps function at an early-stage AI messaging infrastructure startup, working on-site in San... ...production conditions. Collaborate cross-functionally with engineering teams to support development and production workflows....Full timeRemote work$150 per hour
A leading AI data collection platform is seeking AI Trainers - Machine Learning Specialists to assist in training and evaluating advanced AI models. Successful candidates will complete training tasks, provide expert feedback, and help enhance AI's capabilities. This role...Remote jobWork from home$40 per hour
Prolific is seeking AI Trainers with advanced SQL development skills to train and evaluate AI models. The role requires strong attention to detail, ability to work on complex tasks, and a reliable internet connection. Successful candidates will join Prolific as Domain...Remote jobFlexible hours$144k - $187k
...designing and building research infrastructure that integrates AI, developing tools for model development, and maintaining automated validation and... ...S. or Ph.D. in a quantitative field and strong software engineering skills in Python are required. The position offers a...- Armada is seeking a patient, articulate Part-Time Technical Trainer to deliver hands-on training and create demo content for our GPUaaS platform. This role combines technical enablement with customer-facing teaching, traveling to sites worldwide when required, and producing...Part timeRemote workWorldwide
- Mercor, in collaboration with a leading AI lab, seeks experienced FP&A and treasury professionals to translate real planning and cash management work into structured training data for AI that reasons like finance teams. You will design scenarios such as budgets, rolling...
$30 per hour
AI Trainer - Advanced Japanese Fluency Prolific is building the biggest pool of quality human data in the world. Over 35,00... ...looking for advanced Japanese speakers to help train and evaluate cutting‑edge AI models. If you have the necessary experience, you’ll take a...Remote jobHourly paySelf employmentWork from homeFlexible hours- ...Welo Data is seeking Data Labeling Associates in California to evaluate AI outputs and ensure cultural context and safety in Arabic datasets. This role requires professional-level proficiency in Portuguese (Brazil), a bachelor's degree, and at least 2 years of experience...
- Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. You will help evaluate frontier AI coding models through structured technical assessments and focus on realistic data engineering workflows. The role involves reviewing model...
- Obsidian is partnering with an AI research initiative to support a Frontier Code Agents project in San Francisco. You will evaluate frontier AI coding models by performing realistic data engineering tasks, reviewing ETL pipelines, data warehouses, and distributed systems...
$170k - $216k
...of miles of driving data from a diverse set of sensors, enabling engineers like you to (1) develop methods for efficiently and continuously learning from large scale real-world data, to (2) develop models and model training at scale, to (3) analyze real-world behavior and...Full timeRemote work$150k - $200k
...Get AI-powered advice on this job and more exclusive features. Understanding Recruitment provided pay range This range is provided... ...working with a leading AI company, who are looking for a DevOps Engineer to join their ever-expanding Engineering team who can bridge the...Full timeWork at officeImmediate start2 days per week3 days per week$150k - $240k
...DevOps EngineerTitle of Role: DevOps EngineerLocation: San Francisco, on-site or remoteCompany... ...space, particularly within API SDKs and AI-driven solutions. With a strong... ...ElasticSearch, RDS, and Redshift.Collaborate with engineering teams to streamline CI/CD pipelines using...- ...Senior Software Engineer, InfrastructureWe are looking for a Senior Software Engineer, Infrastructure with ideally 5+ years of experience... ...build, release, and runtime foundations at a fast-scaling voice AI company. You'll design and automate deployment pipelines across...
- A technology consulting firm in California is seeking a Senior DevOps Engineer to enhance cloud-based data management platforms. The ideal candidate will have strong AWS and CI/CD skills, alongside expertise in tools like Databricks and Collibra. Duties include defining...Contract work
- ...Senior DevOps EngineerStack AI is building cutting-edge infrastructure at the intersection of AI and enterprise. Our mission is to enable companies... ...AI at scale. As we grow, we're looking for a Senior DevOps Engineer who can design, build, and manage the systems that power...Work at officeImmediate startRemote work
- A leading data and AI company in San Francisco is seeking a Senior Engineer to enhance their Model Serving platform. This role requires expertise in building large-scale distributed systems and collaboration across teams to optimize performance and reliability. Ideal candidates...
$67k - $155.3k
...story. The opportunity As an FSO DevOps Engineer Senior Analyst, you’ll be based in our... ...our team-led and leader-enabled hybrid model. Our expectation is for most people in external... ...capital markets. Enabled by data, AI and advanced technology, EY teams help...Full timeSummer holidayFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to DevOps Engineer - AI Model Evaluator - AI Trainer. Be the first to apply!
- senior devops engineer remote San Francisco, CA
- devops engineer sre San Francisco, CA
- devops engineer San Francisco, CA
- devops aws developer (remote) San Francisco, CA
- devops engineer remote San Francisco, CA
- senior devops cloud engineer San Francisco, CA
- big data devops engineer San Francisco, CA
- senior devops engineer San Francisco, CA
- devops engineer full time San Francisco, CA
- devops cloud engineer San Francisco, CA




