Analytics Data Engineer, Applied Engineering
About the team
The Applied team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses.
We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth.
About the role:
We're seeking a Data Engineer to take the lead in building our data pipelines and core tables for OpenAI. These pipelines are crucial for powering analyses, safety systems that guide business decisions, product growth, and prevent bad actors. If you're passionate about working with data and are eager to create solutions with significant impact, we'd love to hear from you. This role also provides the opportunity to collaborate closely with the researchers behind ChatGPT and help them train new models to deliver to users. As we continue our rapid growth, we value data-driven insights, and your contributions will play a pivotal role in our trajectory. Join us in shaping the future of OpenAI!
In this role, you will:
Design, build and manage our data pipelines, ensuring all user event data is seamlessly integrated into our data warehouse.
Develop canonical datasets to track key product metrics including user growth, engagement, and revenue.
Work collaboratively with various teams, including, Infrastructure, Data Science, Product, Marketing, Finance, and Research to understand their data needs and provide solutions.
Implement robust and fault-tolerant systems for data ingestion and processing.
Participate in data architecture and engineering decisions, bringing your strong experience and knowledge to bear.
Ensure the security, integrity, and compliance of data according to industry and company standards.
You might thrive in this role if you:
Have 3+ years of experience as a data engineer and 8+ years of any software engineering experience(including data engineering).
Proficiency in at least one programming language commonly used within Data Engineering, such as Python, Scala, or Java.
Experience with distributed processing technologies and frameworks, such as Hadoop, Flink and distributed storage systems (e.g., HDFS, S3).
Expertise with any of ETL schedulers such as Airflow, Dagster, Prefect or similar frameworks.
Solid understanding of Spark and ability to write, debug and optimize Spark code.
This role is exclusively based in our San Francisco HQ. We offer relocation assistance to new employees.
About OpenAI
OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
We are an equal opportunity employer and do not discriminate on the basis of race, religion, national origin, gender, sexual orientation, age, veteran status, disability or any other legally protected status.
For US Based Candidates: Pursuant to the San Francisco Fair Chance Ordinance, we will consider qualified applicants with arrest and conviction records.
We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link .
At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
Recommended Jobs
Software Engineer
About us We’re working on Sky: natural-language computing for your Mac. Read more about the team, our values, and our vision at Values We are a highly collaborative organization. We beli…
Sr. Software Engineer
About Blumen Blumen is building the operating system to power the next decade of American Infrastructure. A year from founding, we work with some of the largest renewable power, data centers, teleco…
Spring 2026 FOX Entertainment Internship Program - Los Angeles
OVERVIEW OF THE COMPANY FOX Entertainment With a legacy spanning more than 35 years, FOX Entertainment is one of the world’s most recognizable media brands and a prolific content producer acros…
Senior Data Scientist, LATAM
Job Description Join the team redefining how the world experiences design. Hello, g'day, mabuhay, kia ora, 你好, hallo, vítejte! Thanks for stopping by. We know job hunting can be a little ti…
Software Engineer - Backend
Who We Are At Pave, we're building the industry's leading compensation platform, combining the world's largest real-time compensation dataset with deep expertise in AI and machine learning. Our …
Physical Therapist (PT) for Home Health
This position is for an Independent Contractor to serve the Moorpark Area FeldCare Connects is currently seeking a self-motivated Physical Therapist to deliver premier excellence of care and i…
Superintendent (commercial construction)
Summary Our client is a full-service commercial construction general contractor working on high-quality commercial and residential projects across Southern California — is seeking an experien…
Deal Data Technology & Analytics, Manager Save for Later Remove job
A career in Technology and Data Solutions practice, within Deals M&A Transaction Services, provides the opportunity to help organizations realize the potential of mergers, acquisitions, divestiture…
Finance and Administration Director
Finance and Administration Director Location Hybrid work in Sacramento, CA : POSITION OVERVIEW The full time Finance and Administration Director is a senior-level position that supports the accoun…
Locum CRNA - Nurse Anesthetist
Locum CRNA Opportunity Northern California (Vacaville & Vallejo Area) Assignment Details: Start Date: ASAP 3-month assignment with option to extend Setting: Hospital & Ambulatory Sur…