AI Algo Eng - LLM/VLM Mandarin Required
On-siteSan Francisco, California, United States
Job Summary
Develop production AI systems based on LLMs, VLMs, and multimodal foundation models to enhance semantic understanding, classification, and reasoning across text, images, video, and audio. Design and build scalable Agent systems utilizing ReAct, PlanAct, CodeAct, and multi-agent architectures for task planning, intent understanding, and real-world product integration. Participate in the full model lifecycle including pre-training, supervised fine-tuning, reinforcement learning, and post-training to improve reasoning and multimodal capabilities. Evaluate emerging developments in frontier AI, translate research into scalable production systems, and contribute to technical direction and model evaluation methodologies. Requires Mandarin communication skills for collaboration with China-based teams.
Required Qualifications
- Master's degree or above in Computer Science, Artificial Intelligence, Machine Learning, Mathematics, or related disciplines
- Approximately 3–5+ years of relevant industry experience
- Strong understanding of Large Language Models
- Strong understanding of Vision Language Models
- Strong understanding of Multimodal Foundation Models
- Strong understanding of Agentic AI
- Hands-on experience with Prompt Engineering
- Hands-on experience with Retrieval-Augmented Generation (RAG)
- Hands-on experience with Agent evaluation frameworks
- Hands-on experience with Model evaluation
- Experience with modern Agent architectures, including ReAct
- Experience with modern Agent architectures, including PlanAct
- Experience with modern Agent architectures, including CodeAct
- Experience with modern Agent architectures, including Multi-Agent Systems
- Experience with modern Agent architectures, including Context Engineering
- Experience with modern Agent architectures, including Function Calling
- Experience with modern Agent architectures, including MCP
- Experience with modern Agent architectures, including A2A
- Familiarity with SFT
- Familiarity with RLHF / RL
- Familiarity with Post-Training
- Familiarity with Reward Models
- Familiarity with CodeRL or equivalent reinforcement learning infrastructure
- Strong engineering skills with the ability to translate research ideas into production AI systems
- Professional Mandarin communication skills
Desired Qualifications
- Production experience with Vision Language Models or multimodal foundation models
- Experience building production Agent systems at significant scale
- Background in recommendation systems, search, content understanding, trust & safety, or content moderation
- Experience with large-scale model training or post-training
- Publications at conferences such as NeurIPS, ICLR, ICML, CVPR, ICCV, ACL, EMNLP, or KDD
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.