Zyphra logo
Zyphra
Verified live 1d ago

Research Scientist - Agency and Reasoning

Brief overview

San FranciscoIn-person
MastersOr in progress
4 H-1B approvalsDept. of Labor
Reinforcement learningPost-trainingHuman preference learningLanguage model reasoningSupervised fine-tuningPreference learning (DPO, simPO)Context-length extensionData engineeringSynthetic data generationPyTorchPythonResearch implementationCommunicationCollaboration

About the company

Zyphra is a full stack artificial intelligence company based in Palo Alto, California.

Visa sponsorship history

2 years sponsoring, last filed FY2026

Data powered by U.S. Department of Labor. This does not guarantee sponsorship for this specific role.
4H-1B approved
100%approval rate
$187,574median wage / yr
H-1B Petition ApprovalsVisas USCIS actually granted: the strongest sign the company sponsors.
20251
20263
LCA Certified ApplicationsAn early filing step, not a visa approval: it signals intent, not confirmed sponsorship.
20263
Top sponsored roles
Data Infrastructure Engineer (Member of Technical Staff)Member of Technical Staff (LLM Performance & Optimization)Software Developer (Member of Technical Staff)

Job description

Zyphra is an artificial intelligence company based in San Francisco, California.

The Role:

As a Research Scientist, you will be a core contributor to Zyphra’s Agency and Reasoning Team. You will be involved with performing novel research in reinforcement learning, post-training, and human preference learning, and applying your ideas at scale to our next generation of language models.

What We’re Looking For:

  • Strong research taste and intuition

  • The ability to work through a research project from conception to execution to write-up

  • Strong implementation and prototyping skillset

  • A researcher who can take an idea from conception to experimentation extremely quickly

  • The ability to work well and cooperate with others in a high-paced research setting

  • Curiosity, interest, and joy in understanding intelligence.

Qualifications:

  • Experience and aptitude with reinforcement learning, either in the context of language model reasoning or more classical RL tasks

  • Experience with language-model-supervised fine-tuning and preference-learning methods, such as DPO and simPO.

  • Experience with context-length extension methods

  • A good intuitive ability to understand model behaviors and correct them through iterative fine-tuning

  • Interest in grappling in detail with data and spending significant time involved in data engineering and synthetic data generation

  • Postgraduate degree in a scientific subject (Computer Science, EE/EECS, Mathematics, Physics)

  • Previously published machine learning research in well-respected venues

  • Highly proficient with PyTorch and Python

  • We are excited and able to rapidly learn new fields and implement new ideas

  • Excellent communication and collaboration skills, and can work effectively on both research and engineering implementation at scale

Why Work at Zyphra:

  • We strongly value new and crazy ideas and are very willing to bet big on new ideas

  • We move as quickly as we can; we aim to minimize the bar to impact as low as possible

  • We all enjoy what we do and love discussing AI

Benefits and Perks:

  • Comprehensive medical, dental, vision, and FSA plans

  • Competitive compensation and 401(k)

  • Relocation and immigration support on a case-by-case basis

  • On-site meals prepared by a dedicated culinary team; Thursday Happy Hours

  • In-person team in San Francisco, California with a collaborative, high-energy environment

More jobs like this