Backend Software Engineer (Evals) Support Automation Engineering

OpenAI
San Francisco, CA

About the Team

The Support Automation team at OpenAI scales the organization by applying cutting-edge AI models to real-world challenges, automating and enhancing work across the organization. From customer operations to engineering, we develop an ecosystem of automation products that empower our colleagues and drive impact. We're passionate about crafting products that serve those around us, blending rapid prototyping with a focus on long-term quality and reliability. By creating reusable solutions, we create patterns that can be applied across diverse domains within OpenAI.

TLDR: this team leverages OpenAI technology to improve OpenAI, and you’ll have the opportunity to leverage the full extent of our tech (both public and pre-released) to accomplish this mission.

About the Role

We’re looking for a Backend Software Engineer with experience working in ML/LLM-heavy domains to help to design and build an evals infrastructure that measures the quality of OpenAI’s support automation. This is a deeply technical and highly cross-functional role where you’ll build robust systems and backend services that serve as the foundation for how knowledge is created, accessed, and applied across OpenAI. The role will especially focus on working closely with Data Science and Research partners to design and build evals at scale.


In this role, you will:

  • Design eval pipelines that are reliable, reproducible, and extendable

  • Build the infrastructure for continuous eval monitoring frameworks (regression/drift monitoring, building robust golden datasets) along with feedback loops that ultimately strengthen support automation

  • Design, build, and maintain backend services and APIs to support intelligent automation and knowledge systems

  • Integrate and structure data across internal platforms, transforming it into formats optimized for use by downstream systems and AI workflows.

  • Collaborate closely with data, research, and engineering teams to integrate OpenAI models into high-leverage workflows

  • Own the full development lifecycle of new backend systems and internal platform capabilities

  • Build with scale and maintainability in mind, while rapidly iterating on new ideas

You might be a great fit if you have:

  • 4+ years of backend engineering experience at product-driven companies (excluding internships)

  • Proficiency in backend technologies. Our tech stack includes Python, FastAPI, and Postgres

  • Experience designing and scaling distributed systems, APIs, or data processing pipelines

  • Have experience building AI agents or applications, including designing evals and improving performance through prompting or scaffolding

  • Are familiar with evaluation methods for LLMs and have worked with patterns like multi-agent workflows, tool use, or long context.

  • Experience creating production evals and/or measuring performance of ML/LLM models at scale

  • A pragmatic mindset. You’re comfortable shipping iteratively while building toward a long-term vision

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement .

Qualified applicants with arrest or conviction records will be considered for employment in accordance with applicable law, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.

To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance.

We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link .

At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.

Posted 2025-09-22

Recommended Jobs

Embedded Software Engineer

Neuralink
Fremont, CA

About Neuralink: We are creating devices that enable a bi-directional interface with the brain. These devices allow us to restore movement to the paralyzed, restore sight to the blind, and revolut…

View Details
Posted 2025-09-14

Senior Software Engineer

Rainmaker Systems
Campbell, CA

Job Description We are looking for a few exceptional software engineers to work on our cloud based B2B e-commerce, renewals and subscriptions platform.  As a member of the engineering team…

View Details
Posted 2025-09-14

Lead Software Engineer (Maya)

Scanline Vfx
Los Angeles, CA

As Lead Software Engineer, you would lead a team of engineers to write and maintain the tools necessary to support VFX workflows with a focus on Maya. Our ideal candidate is able to collaborate with …

View Details
Posted 2025-09-22

Production Manager

CRH
Madera, CA

Exempt Oldcastle Infrastructure™, a CRH company, is the leading provider of utility infrastructure solutions for the water, energy, and communications markets throughout North America. We’re…

View Details
Posted 2025-10-31

Staff AI Software Engineer

Komodo Health
San Francisco, CA

We Breathe Life Into Data At Komodo Health, our mission is to reduce the global burden of disease. And we believe that smarter use of data is essential to this mission. That’s why we built the He…

View Details
Posted 2025-10-01

Botany Mentor

National Experienced Workforce Solutions
Solvang, CA

Botany MentorID:F31CAR5-003Location:SolvangProgram:FORESTWage/Hr:$55.00Hours/Week:16Minimum Age:55Duties will be mostly computer-based remote work, with occasional fieldwork within Los Padres National…

View Details
Posted 2025-09-07

Senior/Staff Machine Learning Engineer, Autonomy Validation

Zoox
Foster, CA

Zoox is on an ambitious journey to develop a full-stack autonomous mobility solution for cities and safely deploy a robotaxi service. Zoox’s System Design and Mission Assurance (SDMA) team is a found…

View Details
Posted 2025-10-01

Project Engineer/Assistant PM

Subzero Excavating
Simi Valley, CA

Job Description Job Description A Simi Valley heavy civil/underground utilities contractor is looking for a highly organized, professional, driven, and energetic Office/Project Engineer/Assistant…

View Details
Posted 2025-07-30

Automotive Full-Stack Software Engineer

Gruve
California

About Gruve Gruve is an innovative software services startup dedicated to transforming enterprises to AI powerhouses. We specialize in cybersecurity, customer experience, cloud infrastructure, a…

View Details
Posted 2025-09-22

Graphic Designer (FT)

iKrusher
Baldwin Park, CA

Graphic Designer (FT) Location Baldwin Park, CA : iKrusher is one of the leading brands of vape technology hardware. The HQ is located near Los Angeles, California, in the City of Arcadia, and includ…

View Details
Posted 2025-10-31