Staff AI Engineer, Model Post-Training and Alignment
OKX
Location: APAC
**Staff AI Engineer, Model Post-Training and Alignment at OKX in APAC** **What you'll do** - Lead and execute the full post-training pipeline for large language models (LLMs) - Design and implement advanced training paradigms such as DPO and GRPO - Develop domain-specific data recipes and curation strategies - Build and refine Reward Models to support alignment **What they're looking for** - Bachelor's in Computer Science, AI, Machine Learning, or related fields with 8+ years industry experience - Strong hands-on experience across the full post-training pipeline for large models - Deep familiarity with preference learning and alignment techniques **Details** - Location: APAC - Type: Not specified - Salary: Not specified