Site Reliability Engineer
SRE
Location: San Francisco, CA (5 Days In-Office)
You are the infrastructure expert who enables our rapid product development and guarantees 99.9%+ stability and performance of our clinical AI platform for major health systems. Your focus on operational excellence is directly tied to a patient's access to life-saving treatment.
What We Look for in a Great Engineer
You have the intensity and technical mastery to own mission-critical infrastructure. You hold yourself and others to high standards and thrive in a high-energy, in-office culture where everyone is in it to win it.
Tool Proficiency: You are highly proficient with your tools—you speak command line fluently and have mastered keyboard shortcuts.
Ownership: You thrive on owning complex systems and have a proven track record of scaling mission-critical deployments.
Automation Drive: You love automating things, always finding new ways to increase your own leverage, and defining standards for operational excellence.
Problem Solver: You won't wait for someone else to solve a problem that you're in a position to solve; you are willing to jump into whatever needs to get done.
What You'll Work On (Responsibilities)
As our SRE, you will own the entire production environment and improve the development experience:
Infrastructure Ownership: Design, implement, and maintain the production environment, having previously handled 500+ machine deployments .
Kubernetes Mastery: Own our containerized infrastructure, leveraging deep expertise in Kubernetes and Helm to manage deployment, scaling, and operational health.
CI/CD & Deployment Optimization: Optimize and streamline both the TypeScript and Python/ML deployment pipelines to support high-velocity feature release while maintaining the highest reliability.
DevX Support: Support Developer Experience (DevX) work to streamline developer workflows, enhance tool proficiency, and improve CI/CD systems.
Infrastructure as Code (IaC): Manage and maintain infrastructure definitions using Terraform.
Technical Qualifications & Environment
IaC & Orchestration: Deep, demonstrable experience with Kubernetes, Helm, and Terraform .
Scaling Systems: Proven ability to architect and maintain complex, distributed systems with high-availability requirements.
Deployment Experience: Hands-on experience optimizing deployment pipelines for both application code (TypeScript) and machine learning models (Python/ML). Also PostgreSQL, Redis, Kakfa.
Core Team Member: Excitement about working five days per week in our San Francisco office.
Pay Scale $165,000-250,000
Recommended Jobs
GNC Software Engineer (Falcon)
SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technolog…
Java Developer
Our Client is a leading company in Information Technology (IT) with a solid track record across Mexico and Latin America. They are dedicated to driving the digital transformation and sustainable grow…
Area Sales Manager - North America
Are you ready to represent the world’s leading pure-play MEMS foundry in one of the most dynamic and innovative technology markets in the world? As Area Sales Manager – North America, you will pla…
Personal Injury Attorney
Personal Injury Attorney Location: Torrance, CA Country: United States Salary: $125K-175K Start Date: Description: Job Overview: Are you a dedicated litigator with a passion for…
Caltrans Equipment Operator II
Job Description and Duties Under the supervision of a Caltrans Maintenance Supervisor the Equipment Operator II is responsible for operating and servicing highway maintenance, landscape or constru…
Machine Operator I-ARP
**Accelerate the possible by joining a winning Amcor team that's transforming the packaging industry and improving lives around the world.** At Amcor, we unpack possibility through our innovative and …
Policy and Regulatory Affairs Intern
Policy and Regulatory Affairs at Zoox is responsible for advancing Zoox's policy and critical government affairs priorities through external engagement and advocacy with federal, state, and local ele…