Backend Software Engineer (Evals)
$230k - $385kOpenAI
About the Team
The Support Automation team at OpenAI scales the organization by applying cutting-edge AI models to real-world challenges, automating and enhancing work across the organization. From customer operations to engineering, we develop an ecosystem of automation products that empower our colleagues and drive impact. We're passionate about crafting products that serve those around us, blending rapid prototyping with a focus on long-term quality and reliability. By creating reusable solutions, we create patterns that can be applied across diverse domains within OpenAI.
TLDR: this team leverages OpenAI technology to improve OpenAI, and you'll have the opportunity to leverage the full extent of our tech (both public and pre-released) to accomplish this mission.
About the Role
We're looking for a Backend Software Engineer with experience working in ML/LLM-heavy domains to help to design and build an evals infrastructure that measures the quality of OpenAI's support automation. This is a deeply technical and highly cross-functional role where you'll build robust systems and backend services that serve as the foundation for how knowledge is created, accessed, and applied across OpenAI. The role will especially focus on working closely with Data Science and Research partners to design and build evals at scale.
In this role, you will:
Design eval pipelines that are reliable, reproducible, and extendable
Build the infrastructure for continuous eval monitoring frameworks (regression/drift monitoring, building robust golden datasets) along with feedback loops that ultimately strengthen support automation
Design, build, and maintain backend services and APIs to support intelligent automation and knowledge systems
Integrate and structure data across internal platforms, transforming it into formats optimized for use by downstream systems and AI workflows.
Collaborate closely with data, research, and engineering teams to integrate OpenAI models into high-leverage workflows
Own the full development lifecycle of new backend systems and internal platform capabilities
Build with scale and maintainability in mind, while rapidly iterating on new ideas
You might be a great fit if you have:
4+ years of backend engineering experience at product-driven companies (excluding internships)
Proficiency in backend technologies. Our tech stack includes Python, FastAPI, and Postgres
Experience designing and scaling distributed systems, APIs, or data processing pipelines
Have experience building AI agents or applications, including designing evals and improving performance through prompting or scaffolding
Are familiar with evaluation methods for LLMs and have worked with patterns like multi-agent workflows, tool use, or long context.
Experience creating production evals and/or measuring performance of ML/LLM models at scale
A pragmatic mindset. You're comfortable shipping iteratively while building toward a long-term vision
About OpenAI
OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.
For additional information, please see OpenAI's Affirmative Action and Equal Employment Opportunity Policy Statement.
Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.
To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance.
We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link.
OpenAI Global Applicant Privacy Policy
At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
Compensation Range: $230K - $385K
$175k - $225k
...Senior Backend Engineer In person 5 days/week in San Francisco, Boston, MA, New York. We are looking for a Senior Backend Engineer to... ...the backend systems that power LangChain's observability and evals platform. You will work on the core services that allow developers...SuggestedWork at officeFlexible hours$170k - $195k
...agents ubiquitous. We provide the agent engineering platform and open source frameworks developers... ...London. We are looking for a Senior Backend Engineer to join us. In this role you... ...power LangChain’s observability and evals platform. You will work on the core services...SuggestedWorldwideFlexible hours$170k - $195k
...A tech company specializing in AI is looking for a Senior Backend Engineer to build backend systems for their observability and evaluation platform. This role requires over 5 years of experience in backend engineering and proficiency in languages like Python or Go. The...Suggested$275k - $325k
...About Hazel: Hazel.ai is building the AI engine for wealth management that unlocks 10x growth... .... You'll work shoulder-to-shoulder with backend engineers, product managers, and a... ...signals. Your impact: Design and build Hazel's evals platform end-to-end – online scoring,...SuggestedWork at officeImmediate start$140k - $175k
...ubiquitous. We provide the agent engineering platform and open source... ...commercial observability and evals platform product. In this role... ...~2+ years of experience in software engineering working on complex... ...experience with Go or Python on the backend and React + Typescript on the...SuggestedWorldwideFlexible hours$255k - $405k
...organization. From customer operations to engineering, we develop an ecosystem of automation... ...About the Role We’re looking for a Backend Software Engineer with experience working in ML... ...domains to help to design and build an evals infrastructure that measures the...InternshipWork at officeLocal areaRelocation packageFlexible hours- ...next generation of built-world software will not just organize work.... ...We are looking for an AI engineer, core to build the infrastructure... ...have deep experience with backend systems, distributed systems,... ...calling, structured outputs, RAG, evals, tracing, or agent frameworks...
- ...Plaid Inc is seeking a Software Engineer in San Francisco to build and maintain scalable backend systems and APIs. This role involves collaborating closely with product managers and designers, writing clean and efficient code, and contributing to quality assurance through...
- ...Finch in San Francisco is seeking a Backend Developer to design and maintain services that power their data presentation layer. You'll contribute to exciting product initiatives and work closely with cross-functional teams. The ideal candidate has over 4 years of backend...Work at officeRemote work2 days per week
- ...Perplexity is seeking a Backend Software Engineer to join their San Francisco-based team focused on revolutionizing internet search and interaction. The role involves designing, implementing, and scaling backend systems that support web, mobile, and browser products, using...
$175k - $300k
...great founding team in SF. Shivaal Roy (CTO) was a founding engineer at Glean ($100M+ ARR today) and managed the Assistant and Search... ...~ Track record of extraordinary ability ~3 - 8 years of software engineering experience ~ Excitement to work in person 4 days...- ...to our 2+ year lead and regulatory advantages), we are setting out to become a $100B+ company. We're seeking a talented Backend Software Engineer to join our innovative team and play a crucial role in developing our AI-driven healthcare platform. As a key member of...Full timeWork at officeMonday to Friday
- ...gets bigger. What you'll do Build and ship scalable backend/platform systems Develop LLM agents that integrate with... ...challenges What we're looking for Strong backend and platform engineering experience (3+ years) Ability to work independently, learn...Work at officeFlexible hours
$180k - $280k
...coding, marketing is next. The Engineering Challenge Fully... ...job. Work primarily across backend systems, data pipelines, and... ...observability, and LLM agents, evals, tool-use systems, retrieval/... ...enough to encode judgment into software. You are comfortable working...Work at officeRelocation packageShift work- ...About the Team Core backend engineering is central to delivering reliable, scalable, and performant AI-driven products. Our Backend teams span... ...’s global footprint. About the Role We’re hiring Backend Software Engineers to design and implement the services and infrastructure...
- ...Tiger is a leading software company. Example org allows real-time collaboration on important example workflows. Founded in 2012 we have over 10,000 customers worldwide and are backed by fantastic investors such as Example Capital. Example has raised its Series C and is...Work at officeWorldwideFlexible hours
$170k - $240k
...team is building the autonomous backend that turns healthcare's most... ...from prompt design and evals through production infrastructure... ...looking for a Senior Backend Engineer who takes ownership end-to-end... ...Work across the entire software stack Work with a stack that...Full timeWork at officeImmediate start- ...About the Team The ChatGPT team works across research, engineering, product, and design to bring OpenAI’s technology to the world. We seek... ...products. About the Role We are looking for an experienced backend engineer to join our new ChatGPT Growth team to spearhead high...
$255k - $405k
...learning from every exchange, and finding novel ways to show the value of our technology. About the Role We\'re looking for backend software engineers with a product mindset to join the GTM Innovation team. You\'ll help OpenAI meet the world at scale. You\'ll partner...- ...AWS for AI models"-not data or raw compute, but a full-stack backend for fine-tuning, reinforcement learning, inference, and long-... ...they make your AI system better. They are hiring a Backend Software Engineer (ML Infrastructure) to help design, build, and scale the core...
- ...Senior Backend Engineer At Commure, we're building the AI Operating System for healthcare,... ...of AI features from prompt design and evals through production infrastructure, and... ...and decisions Work across the entire software stack Work with a stack that includes...Full timeImmediate start
$300k
...intelligence. Our team includes world-class AI researchers and engineers from top universities and tech companies, working together... ...with research-grade ambition. We’re looking for exceptional Backend Software Engineers to help architect, build, and scale the...Remote work- ...global scale. Our team partners closely with researchers, product engineers, designers, and platform teams to bring state-of-the-art image... ...can do. About the Role We are looking for an experienced Backend Engineer to join the Image Generation team and help build the systems...Worldwide
$130k - $400k
...Backend Software Engineer Title of Role: Backend Software Engineer Location: San Francisco, onsite Company Stage of Funding: Series C - Software Development Office Type: Onsite Salary: $130K-$400K Company Description We're representing a dynamic...Work at officeRelocation package- ...019 to upgrade manufacturing. We are engineers with deep experience across the product... ...About the role: On Lumafield’s Backend team, you work on the core of our cloud... ...dataset processing using numpy ~ Strong software engineering fundamentals including git,...Full timeWork at officeWork visaFlexible hours
$125k - $150k
...in 2019 to upgrade manufacturing. We are engineers with deep experience across the product... ...Francisco, CA. About the role: On Lumafield’s Backend team, you work on the core of our cloud... ...dataset processing using numpy Strong software engineering fundamentals including git,...Work at officeWork visaFlexible hours$180k - $235k
...experience on the streaming basics, while cooking up next generation multi-screen and multi-user playback experiences. Senior Backend Software Engineer (Video Engineering) Video is at the core of Philo and what customers engage with the most: watching their favorite shows....Full timeFor contractorsWork at officeImmediate startRemote workHome officeFlexible hours3 days per week- ...Rippling, based in San Francisco, is looking for a Software Engineer Intern for Winter 2027. In this role, you’ll join a team to develop robust software solutions, implement updates, and tackle complex challenges. You will gain valuable experience akin to full-time engineers...Full timeInternship
- ...Blockchaincapital.com is seeking a Founding Senior Backend Engineer in San Francisco to lead the development of the backend infrastructure for the Button protocol. The ideal candidate should have significant experience in backend development with modern programming languages...
$170k - $220k
...Merge API Inc. is looking for backend and full stack engineers to join their San Francisco teams, working on innovative AI and integration solutions. Ideal candidates will have 3‑7+ years of software engineering experience, preferably in startup environments. The role...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Backend Software Engineer (Evals). Be the first to apply!
- entry level back-end developer San Francisco, CA
- back-end developer San Francisco, CA
- lead backend developer San Francisco, CA
- backend software engineer San Francisco, CA
- remote back end developer San Francisco, CA
- backend python developer San Francisco, CA
- senior backend developer San Francisco, CA
- backend intern San Francisco, CA
- id software San Francisco, CA
- android software developer San Francisco, CA




