PhD Research Intern, Physical AI in Perception
About Zoox
Zoox is transforming mobility with fully autonomous, electric vehicles designed from the ground up for a driverless future. Our mission is to make transportation safer, more sustainable, and accessible to everyone. At Zoox, innovation, collaboration, and a bold vision for the future drive everything we do.
About Our Internship Program
Zoox’s internship program offers hands-on experience with cutting-edge technology, mentorship from some of the industry’s brightest minds, and the opportunity to make meaningful contributions to real projects. We seek interns who demonstrate strong academic performance, engagement beyond the classroom, intellectual curiosity, and a genuine interest in Zoox’s mission.
Project Overview
This internship is part of the Perception Semantics team, focused on advancing on-robot AI systems that enable machines to understand and interact with the physical AI world. You’ll work on cutting-edge problems in vision-language-action (VLA) modeling, world modeling, spatial reasoning, and mapping, contributing to both research and real-world deployment.
Projects are open-ended and research-driven, giving you the opportunity to explore new ideas, develop novel approaches, and evaluate them in realistic settings. This role is ideal for Ph.D. students interested in pushing the boundaries of computer vision and embodied AI while seeing their work translate into real-world impact.
Qualifications:
- Currently enrolled in the Ph.D program in Computer Science, Electrical/Computer Engineering, or related field, with the specialization in the CV/NLP/ML
- Experience in multi-modal modeling (vision, language, or planning), with deep understanding of Vision Language Model, vision foundation model, flow-matching, temporal modeling, and reinforcement learning techniques
- Strong proficiency in PyTorch and modern transformer-based model design
Bonus Qualifications:
- Publication records in top-tier AI conferences (CVPR, ICCV, ECCV, NeurIPS, ICLR, ICML, etc)
- Prior experience building foundation or end-to-end driving models for autonomous driving or robotics, or working deeply on LLM/VLM architectures (e.g., ViT, Flamingo, BEVFormer, RT-2, or GRPO-style policies)
- Knowledge of RLHF/DPO/GRPO, trajectory prediction for safety
Requirements:
- Currently working towards a Ph.D in a relevant engineering program
- Good academic standing
- Able to commit to a 12-week internship during one of the following summer 2026 cohorts: May 18th - August 7th, OR May 26th - August 14th, OR June 15th - September 4th
- At least one previous industry internship, co-op, or project completed in a relevant area
- Ability to relocate to the Bay Area, California (or Boston, Massachusetts) for the duration of the internship
- Interns at Zoox may not use any proprietary information they are working on as part of their thesis, any published work with their university, or to be distributed to anyone outside of Zoox
Compensation:
The monthly salary for this position is $9,500. Compensation will vary based on geographic location. Additional benefits may include medical insurance, and a housing stipend (relocation assistance will be offered based on eligibility).
About Zoox
Zoox is developing the first ground-up, fully autonomous vehicle fleet and the supporting ecosystem required to bring this technology to market. Sitting at the intersection of robotics, machine learning, and design, Zoox aims to provide the next generation of mobility-as-a-service in urban environments. We’re looking for top talent that shares our passion and wants to be part of a fast-moving and highly execution-oriented team.
Accommodations
If you need an accommodation to participate in the application or interview process please reach out to [email protected] or your assigned recruiter.
A Final Note:
You do not need to match every listed expectation to apply for this position. Here at Zoox, we know that diverse perspectives foster the innovation we need to be successful, and we are committed to building a team that encompasses a variety of backgrounds, experiences, and skills.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
Recommended Jobs
Insurance Partner Operations Lead
Come join VisitorsCoverage, one of Silicon Valley's most successful InsurTech companies, certified as a Great Place to Work ®! We are looking for a Partner Operations Lead to join our growing te…
217354 - Nondestructive Evaluation Technical Analyst 5
Chipton-Ross is seeking an Nondestructive Evaluation Technical Analyst 5 for a contract opportunity in Berkeley, MO. 100% onsite BASIC QUALIFICATIONS (REQUIRED SKILLS/EXPERIENCE): o Ability to ob…
Electronic Assembler
Seeking immediate new hires! Our Production team is hiring Assemblers for Metal Finishing/Masking/Painting jobs that have just opened due to increased production needs. This is an entry-level ro…
Work From Home Opportunity Entry Level (Work from Home) No Experience Needed- Full Training Provided
100% Remote – No Experience Needed – Start This Week! Company: Globe Life AO Employment Type: Full-Time / Part-Time Location: Remote - United States Start Your Career Work-From-Home Oppor…
Academic Instructor (Credential Teacher - Balboa High School)
JOB ANNOUNCEMENT The Community Youth Center of San Francisco (CYC) provides the youth of our city a sense of belonging and vital tools and experiences to succeed in life. Our services include acad…
Server - Restaurant 7
Join Our Team as a Server! As a Server, you’re the dining room’s MVP! You’ll wow our guests by juggling orders with a smile, chatting up the menu like a pro, and turning every visit into a memorable …
Senior Product Manager, Clinical Execution & Team Catalyst
About Enara Enara is a world renowned obesity and medical weight loss start-up, based in Silicon Valley, pioneering the use of data, digital, and clinical treatments to provide personalized plans …
Director, Catalog Visual Creative - Santa Monica, 90404
Director, Catalog Visual Creative - Santa Monica, 90404, United States of America Director, Catalog Visual Creative Interscope Geffen A&M (“IGA”) How We LEAD The Director, Catalog Visual Cr…
Occupational Therapist (Clinical Specialist) - Physical Rehabilitation
Summary This position is eligible for the Education Debt Reduction Program (EDRP) - a student loan payment reimbursement program. You must meet specific eligibility requirements per VHA policy and…
Lead Supervisor
Coach is seeking a Lead Supervisor in San Diego to oversee retail operations, lead a team of sales associates, and implement marketing strategies. The role requires a minimum of 3 years of supervisory…