Research Scientist - Agency and Reasoning

Zyphra
Palo Alto, CA

Job Description

Job Description

Zyphra is an artificial intelligence company based in Palo Alto, California.

The Role:

As a Research Scientist , you will be a core contributor to Zyphra’s Agency and Reasoning Team. You will be involved with performing novel research in reinforcement learning, post-training, and human preference learning, and applying your ideas at scale to our next generation of language models.

What We’re Looking For:
  • Strong research taste and intuition

  • The ability to work through a research project from conception to execution to write-up

  • Strong implementation and prototyping skillset

  • A researcher who can take an idea from conception to experimentation extremely quickly

  • The ability to work well and cooperate with others in a high-paced research setting

  • Curiosity, interest, and joy in understanding intelligence.

Qualifications:
  • Experience and aptitude with reinforcement learning, either in the context of language model reasoning or more classical RL tasks

  • Experience with language model supervised finetuning and preference learning methods such as DPO, simPO, etc.

  • Experience with context-length extension methods

  • A good intuitive ability to understand model behaviors and correct them through iterative fine-tuning

  • Interest in grappling in detail with data and spending significant time involved in data engineering and synthetic data generation

  • Postgraduate degree in a scientific subject (Computer Science, EE/EECS, Mathematics, Physics)

  • Previously published machine learning research in well-respected venues

  • Highly proficient with PyTorch and Python

  • We are excited and able to rapidly learn new fields and implement new ideas

  • Excellent communication and collaboration skills, and can work effectively on both research and engineering implementation at scale

Why Work at Zyphra:
  • We strongly value new and crazy ideas and are very willing to bet big on new ideas

  • We move as quickly as we can; we aim to minimize the bar to impact as low as possible

  • We all enjoy what we do and love discussing AI

Benefits and Perks:
  • Comprehensive medical, dental, vision, and FSA plans

  • Competitive compensation and 401(k)

  • Relocation and immigration support on a case-by-case basis

  • On-site meals prepared by a dedicated culinary team; Thursday Happy Hours

  • In-person team in Palo Alto, CA, with a collaborative, high-energy environment

Posted 2025-07-30

Recommended Jobs

Full Time Primary Care Physician Job Los Angeles, CA

The Inline Group The Inline Group
Los Angeles, CA

The Inline Group - Full Time Employed New Graduates Average Patients seen: 20 Call Schedule: None Compensation: - $220,000 - $250,000 based on experience- Some productivity bonus pr…

View Details
Posted 2025-09-10

Senior Forensic Auditor

HireNow Medical Solutions
Westwood, CA

Job Description Job Description Job Title: Senior Forensic Auditor Location: Westwood, CA (Hybrid Schedule Available) Employment Type: Full-Time | Direct Hire Compensation: $80,000 - …

View Details
Posted 2025-07-30

Full Stack Developer

A Society Group, Inc.
San Francisco, CA

We are dedicated to creating innovative solutions that drive our business forward. Our team is passionate about technology, collaboration, and delivering exceptional results. We are looking for a tal…

View Details
Posted 2025-09-14

Senior Fullstack Software Engineer, Benefits

Rippling
San Francisco, CA

About Rippling Rippling gives businesses one place to run HR, IT, and Finance. It brings together all of the workforce systems that are normally scattered across a company, like payroll, expenses,…

View Details
Posted 2025-09-14

Compliance Officer (Internal audit) - (Bilingual Korean/English)

SBT Global, Inc.
Los Angeles, CA

Job Description Job Description Company Description English/Korean Bilingual Custom clearance compliance Officer will be responsible for developing, managing and monitoring compliance po…

View Details
Posted 2025-07-30

Line Cook

Ridge Creek Golf Club
Dinuba, CA

Property Description: Ridge Creek Dinuba Golf Club sits in the heart of the Central Valley near Fresno, California. It boasts an award-winning par 72 Championship John Fought designed golf course. Th…

View Details
Posted 2025-09-10

Full-Stack Software Engineer (Jr/Mid level)

Teambridge
San Francisco, CA

About Teambridge: More than 60% of workers in the US (and 70% of workers in the world) are paid hourly, and the businesses that employ them each have their own unique processes and workflows whe…

View Details
Posted 2025-09-14

Acoustic Engineering Co-Op (Winter/Spring Term)

Skyworks
Irvine, CA

If you are looking for a challenging and exciting career in the world of technology, then look no further. Skyworks is an innovator of high-performance analog semiconductors whose solutions are power…

View Details
Posted 2025-08-07

Senior Frontend Engineer, AI

Airwallex
San Francisco, CA

About Airwallex Airwallex is the only unified payments and financial platform for global businesses. Powered by our unique combination of proprietary infrastructure and software, we empower over 150…

View Details
Posted 2025-09-12