Building face-to-face AI interaction that feels genuinely human.
Member of Technical Staff — RL Research
Current conversational agents still fall flat the moment eye contact breaks or subtext surfaces. As an RL Researcher, you will architect reward models and training loops that teach neural networks how to navigate the subtle, unspoken cadence of human dialogue. Your work directly bridges the gap between rigid text generation and fluid, in-person conversation, operationalizing reinforcement learning to make AI feel present rather than scripted.
What they're looking for
- Recently completed or即将 completing a PhD in machine learning, reinforcement learning, or a related quantitative field.
- Proficiency translating research into practice using Python and PyTorch.
- Deep familiarity with RL algorithms, reward modeling, and alignment techniques.
- Ability to work on-site in Seattle alongside a focused, tight-knit research team.
- 0 to 2 years of post-graduate or industry research experience.