Senior Product Manager, Multilingual AI and Evals

OKX · Singapore
full-time senior Posted 8 hours ago
Apply Now Stand out: build a proof-of-work pitch →

Free GitHub-based preview. Direct apply stays one click away.

Get weekly job alerts like this →

Hiring for this role?

AI Market Demand Pack · $29 one-time

Compare this role's skills with the full AI hiring market. Get ranked demand, salary bands, leading companies, public source URLs, and a decision brief.

See the live sample →

About this role

OKX will be prioritising applicants who have a current right to work in Singapore, and do not require OKX's sponsorship of a visa. Who We Are At OKX, we believe that the future will be reshaped by crypto, and ultimately contribute to every individual's freedom.   OKX is a leading crypto exchange, and the developer of OKX Wallet, giving millions access to crypto trading and decentralized crypto applications (dApps). OKX is also a trusted brand by hundreds of large institutions seeking access to crypto markets. We are safe and reliable, backed by our Proof of Reserves.    Across our multiple offices globally, we are united by our core principles: We Before Me , Do the Right Thing , and Get Things Done . These shared values drive our culture, shape our processes, and foster a friendly, rewarding, and diverse environment for every OK-er. OKX is part of OKG, a group that brings the value of Blockchain to users around the world, through our leading products OKX, OKX Wallet, OKLink and more.   About The Opportunity We are building an AI-native localization stack: agentic translation workflows that already publish content with minimal human touch, and a quality system that keeps that automation trustworthy at scale. As automation grows, quality assurance becomes the product. We are hiring a Senior PM to own QA and Evals for our multilingual AI agents — the evaluation, testing, and feedback infrastructure that decides whether an AI output is good enough to auto-publish, catches localization defects inside the live product, and continuously feeds signal back to improve our agents. This role sits at the intersection of platform/product building, AI agent evaluation, and global product delivery. It requires sharp judgement and instinct to turn fuzzy, subjective quality problems into measurable, automatable systems, as well as strong product execution to ship the tooling that runs them at scale. What You’ll Be Doing Multilingual AI Quality Evaluation Enhance the quality gate of the agentic pipeline — the evaluation agent that decides whether a translation is good enough to auto-publish. Own the trade-off between automation coverage and risk across content tiers. When the gate gets something wrong systematically rather than as a one-off, diagnose the root cause and drive the fix back into the agent. AI Driven Localization Testing Build an AI-driven product testing tool that automatically detects localization defects that translation review cannot catch — truncation, layout breakage, hardcoded strings, unlocalized images, wrong number/date/currency formats. Ship the MVP that scans mobile/web apps to accelerate localization auditing, then drive the long-term vision of integrating localization tests into internal product testing infrastructure, so they run automatically before features go live. Evaluation & Annotation Infrastructure Build the backbone that makes quality measurable and improvable: golden datasets, annotation tooling with error classification, automated evaluation, a metrics/dashboard layer, and the feedback loop that turns human evaluations into agent fine-tuning. Own annotation workflow integration so linguists can annotate in a standardized, real-time way. What We Look For In You   Bachelor's degree or higher with 3+ years in product management. Hands-on experience building or evaluating AI agent systems (not just using AI tools) may substitute for part of the PM tenure requirement. Platform-building experience: You've built content platforms, testing platforms, or SaaS products — you know how to take a fuzzy internal workflow and turn it into a system others rely on. Experience with CMSes, testing infrastructure, or internal tooling counts. AI agent / evaluation: You've worked hands-on with AI agents, agent harnesses, and evals — especially for subjective, hard-to-measure quality problems (not just accuracy against a clean label). You can design eval harnesses and golden datasets, and read failure modes (hallucination, prompt drift, edge-case regressions) well enough to know whether a bad output is a prompt problem, a data problem, or a model problem. Multi-market shipping experience: You've shipped products across multiple markets or languages — you understand what breaks when a single-market assumption meets a global user base, and you're comfortable working cross-culturally with distributed teams. Excellent cross-functional leadership across engineering, AI teams, linguists, design, and external partners Nice-To-Haves Localization/translation domain background — terminology & glossary management, TMS tools (Phrase, Smartling), internationalization (ICU/CLDR) Bilingual or multilingual (Chinese and/or European languages) Crypto, fintech, or other high-compliance / regulated domain experience Perks & Benefits  Competitive total compensation package L&D programs and education

Similar Jobs

Related searches:

On-site Jobs Senior Jobs On-site Senior Jobs Senior Machine LearningSenior Fintech & Payments AISenior AI ResearchSenior AI Agents & RAGSenior Healthcare AISenior Generative AI AI Jobs in Singapore Machine Learning in SingaporeFintech & Payments AI in SingaporeAI Research in SingaporeAI Agents & RAG in SingaporeHealthcare AI in SingaporeGenerative AI in Singapore healthcarefine-tuningpaymentsagentsevaluation

Get jobs like this delivered weekly

Free AI jobs newsletter. No spam.