Senior Product Manager, Agent QA & Localization Quality
Who We Are
At OKX, we believe that the future will be reshaped by crypto, and ultimately contribute to every individual's freedom.
OKX is a leading crypto exchange, and the developer of OKX Wallet, giving millions access to crypto trading and decentralized crypto applications (dApps). OKX is also a trusted brand by hundreds of large institutions seeking access to crypto markets. We are safe and reliable, backed by our Proof of Reserves.
Across our multiple offices globally, we are united by our core principles: We Before Me, Do the Right Thing, and Get Things Done. These shared values drive our culture, shape our processes, and foster a friendly, rewarding, and diverse environment for every OK-er.
About the Opportunity
We are building an AI-native localization stack: agentic translation workflows that already publish content with minimal human touch, and a quality system that keeps that automation trustworthy at scale. As automation grows, quality assurance becomes the product. We are hiring a Senior PM to own Agent QA — the evaluation, testing, and feedback infrastructure that decides whether an AI translation is good enough to auto-publish, catches localization defects inside the live product, and continuously feeds signal back to improve our agents.
This is a bridge role for someone who is genuinely strong in both: you bring real localization/translation domain depth (you know what makes a translation wrong, what breaks in-product, how terminology and locale rules work) and the AI product instinct to turn that judgment into agents, metrics, and automated tooling. You will define what "good" means, build the tooling to measure it automatically, and make our human experts dramatically more leveraged.
What You’ll Be Doing
AI Localization Testing Tool
Build an AI-driven product testing tool that automatically detects localization defects that translation review cannot catch — truncation, layout breakage, hardcoded strings, unlocalized images, wrong number/date/currency formats. Ship the MVP that scans mobile/web apps to accelerate localization auditing, then drive the long-term vision of integrating localization tests into the internal product testing infrastructure, so they run automatically before features go live.
Agentic Translation Quality Evaluation
Enhance the quality gate of the agentic pipeline — the quality evaluation agent that decides whether a translation is good enough to auto-publish. Own the trade-off between automation coverage and risk across content tiers. When the gate gets something wrong systematically rather than as a one-off, diagnose the root cause, and drive the fix back into the agent.
Evaluation & Annotation Infrastructure
Build the backbone that makes quality measurable and improvable: golden datasets, annotation tooling with error classification, automated evaluation, a metrics/dashboard layer, and the RLHF feedback loop that turns human evaluations into agent fine-tuning. Own annotation workflow integration (e.g. with internal AI infra) so linguists can annotate in a standardized, real-time way.
What We Look For In You
-
5+ years in product / program management, including hands-on localization experience — e.g. terminology & glossary management, AI translation workflows, Translation Management Systems such as Phrase and Smartling, internationalization (ICU/CLDR), and international product launches. You have owned localization quality in practice, not just adjacent to it.
-
AND demonstrated experience shipping AI / ML or evaluation-driven products (or internal tooling) — you can partner deeply with AI teams, not just consume their output.
-
Hands-on fluency with LLM evaluation and eval-driven development: designing eval harnesses/golden datasets, prompt and QE design, precision/recall tuning against gold sets, and reading failure modes (hallucination, prompt drift, edge-case regressions) to know whether a bad output is a prompt problem, a data problem, or a model problem.
-
Strong data fluency — define metric frameworks, spec dashboards, and reason about precision/recall trade-offs (e.g. false-flag vs miss rate). Ability to define metrics for fuzzy, subjective quality problems and turn them into measurable, automatable systems.
-
Excellent cross-functional leadership across engineering, AI teams, linguists, design, and external partners; comfortable working across multiple markets/languages in a multicultural environment.
Nice to Haves
-
Bilingual or multilingual (European languages and/or Chinese); experience shipping products in multi-market, multi-language environments.
-
Experience building content systems, or working with CMSes like Contentful or Webflow. Familiarity with design tooling, such as Figma.
-
Crypto, fintech, or other high-compliance / regulated domain experience.
Perks & Benefits
-
Competitive total compensation package
-
L&D programs and Education subsidy for employees' growth and development
-
Various team building programs and company events
-
Wellness and meal allowances
-
Comprehensive healthcare schemes for employees and dependants
-
More that we love to tell you along the process!
#LI-WWW
#LI-ONSITE
#LI-ONSITE
Notice:
All official OKX vacancies are published on this website. While roles may appear on selected third-party platforms from time to time, information on other sites may be inaccurate or outdated. If in doubt, please apply directly through our official careers website.
Information collected and processed as part of the recruitment process of any job application you choose to submit is subject to OKX's Candidate Privacy Notice.