Senior Staff Research Engineer – Reinforcement Learning for AI Agents
XPENG
Applying takes a free account: you'll come right back to this job to finish with your profile.
Need a reasonable accommodation to apply or interview? Contact us.
JobMinglr uses automated technology to recommend jobs based on profile information and job preferences. Match Score does not determine eligibility for a position, prevent a user from viewing or applying to a job, or make hiring decisions on behalf of an employer.
Description
Key Responsibilities:
- Reinforcement learning methods for LLM-driven agents and decision systems.
- Policy optimization for long-horizon reasoning and planning.
- Learning from human or AI feedback (RLHF / RLAIF).
- Agent training pipelines built on top of our agent infrastructure platform.
- Evaluation and benchmarking systems for agent capabilities.
- Learning loops that integrate real-world and simulation data.
- Contribute to AI systems that continuously improve after deployment.
Basic Qualifications
- MS or PhD in Computer Science, AI, Machine Learning, Robotics, or a related field.
- Strong background in reinforcement learning or machine learning.
- Experience implementing RL algorithms such as PPO, Actor-Critic, or policy gradient methods.
- Strong programming skills in Python with PyTorch or JAX.
- Experience building ML training systems or infrastructure.
Preferred Qualifications
- Experience with RLHF or preference learning.
- Experience with LLM agents or tool-using AI systems.
- Multi-agent systems or long-horizon planning.
- Simulation environments for RL.
- Publications in NeurIPS, ICML, ICLR, ACL, or related venues.
- A fun, supportive and engaging environment.
- Opportunity to make significant impact on transportation revolution by the means of advancing autonomous driving.
- Opportunity to work on cutting edge technologies with the top talent in the field.
- Competitive compensation package.
- Snacks, lunches and fun activities.
Pay for this role
The employer didn't post a pay range for this role. That usually means pay is set in negotiation, which favors whoever arrives with numbers. Check ranges on comparable Senior Staff Research Engineer – Reinforcement Learning for AI Agents postings in Santa Clara, and analyze any offer before you accept it.
Before you apply, worth reading
How JobMinglr reads this job
Every listing here is scored against your profile before you apply: skills overlap, experience level, location and work arrangement, each weighted and explained. You see the score and the reasons, not just a list. How the matching works.