A tiny, well-funded team pushing frontier intelligence to run on any device.
Member of ML Technical Staff
Referral bonus eligible. Know someone who would be great for this role? Refer them to a Fluency Digital recruiter. If they're placed, you may be eligible for a referral bonus of up to $5,000 (terms apply). Refer someone for this role
You are the engineer who obsesses over latency curves and knows exactly how many FLOPs fit on a mobile NPU. Here, you will rewrite the rules of on-device inference by merging cutting-edge research with brutalist C++ optimization. Forget abstract prototypes; your PyTorch models must survive the transition to production hardware without losing a single beat of accuracy.
What they're looking for
- 1+ years of experience deploying machine learning models in production environments
- Deep proficiency in PyTorch alongside expert-level C or C++ coding skills
- Proven track record of optimizing neural networks for edge or mobile hardware
- Ability to thrive in a microscopic, high-impact engineering collective
- Relocation readiness for an on-site role in San Francisco