We are building an AI-native localization stack: agentic translation workflows that already publish content with minimal human touch, and a quality system that keeps that automation trustworthy at scale. As automation grows, quality assurance becomes the product. We are hiring a Senior PM to own QA and Evals for our multilingual AI agents — the evaluation, testing, and feedback infrastructure that decides whether an AI output is good enough to auto-publish, catches localization defects inside the live product, and continuously feeds signal back to improve our agents.
This role sits at the intersection of platform/product building, AI agent evaluation, and global product delivery. It requires sharp judgement and instinct to turn fuzzy, subjective quality problems into measurable, automatable systems, as well as strong product execution to ship the tooling that runs them at scale.