The vacancy is well-defined with clear responsibilities and requirements, but lacks compensation specifics and tech stack details.
Check Match — Just drop your CV
See your fit for Senior Product Manager, Multilingual AI and Evals in seconds.
Overview
Join OKX as a Senior Product Manager to lead QA and evaluations for multilingual AI agents, enhancing automation and quality assurance in crypto applications. At OKX, we believe that the future will be reshaped by crypto, and ultimately contribute to every individual's freedom. OKX is a leading crypto exchange, and the developer of OKX Wallet, giving millions access to crypto trading and decentralized crypto applications (dApps). OKX is also a trusted brand by hundreds of large institutions seeking access to crypto markets. We are safe and reliable, backed by our Proof of Reserves. Across our multiple offices globally, we are united by our core principles: *We Before Me*, *Do the Right Thing*, and *Get Things Done*. These shared values drive our culture, shape our processes, and foster a friendly, rewarding, and diverse environment for every OK-er. OKX is part of OKG, a group that brings the value of Blockchain to users around the world, through our leading products OKX, OKX Wallet, OKLink and more.
Multilingual AI Quality Evaluation
- •Enhance the quality gate of the agentic pipeline — the evaluation agent that decides whether a translation is good enough to auto-publish.
- •Own the trade-off between automation coverage and risk across content tiers. When the gate gets something wrong systematically rather than as a one-off, diagnose the root cause and drive the fix back into the agent.
AI Driven Localization Testing
- •Build an AI-driven product testing tool that automatically detects localization defects that translation review cannot catch — truncation, layout breakage, hardcoded strings, unlocalized images, wrong number/date/currency formats.
- •Ship the MVP that scans mobile/web apps to accelerate localization auditing, then drive the long-term vision of integrating localization tests into internal product testing infrastructure, so they run automatically before features go live.
Evaluation & Annotation Infrastructure
- •Build the backbone that makes quality measurable and improvable: golden datasets, annotation tooling with error classification, automated evaluation, a metrics/dashboard layer, and the feedback loop that turns human evaluations into agent fine-tuning.
- •Own annotation workflow integration so linguists can annotate in a standardized, real-time way.
Perks & Benefits
- •Competitive total compensation package
- •L&D programs and education subsidy for employees' growth and development
- •Various team building programs and company events
- •Wellness and meal allowances
- •Comprehensive healthcare schemes for employees and dependants
- •More that we love to tell you along the process!
What We Look For In You
- •Bachelor's degree or higher with 3+ years in product management. Hands-on experience building or evaluating AI agent systems (not just using AI tools) may substitute for part of the PM tenure requirement.
- •Platform-building experience: You've built content platforms, testing platforms, or SaaS products — you know how to take a fuzzy internal workflow and turn it into a system others rely on. Experience with CMSes, testing infrastructure, or internal tooling counts.
- •AI agent / evaluation: You've worked hands-on with AI agents, agent harnesses, and evals — especially for subjective, hard-to-measure quality problems (not just accuracy against a clean label). You can design eval harnesses and golden datasets, and read failure modes (hallucination, prompt drift, edge-case regressions) well enough to know whether a bad output is a prompt problem, a data problem, or a model problem.
- •Multi-market shipping experience: You've shipped products across multiple markets or languages — you understand what breaks when a single-market assumption meets a global user base, and you're comfortable working cross-culturally with distributed teams.
- •Excellent cross-functional leadership across engineering, AI teams, linguists, design, and external partners.