xai
Member of Technical Staff - Post-Training and RL
At a Glance
- Location
- Palo Alto, California, United States
- Compensation
- s. COMPENSATION AND BENEFITS: $180,000 - $600,000 USD Base salary is just one p
- Posted
- 2026-04-28T23:18:00-04:00
Requirements
You believe truth-seeking AI is the most important and challenging problem.
You are obsessed about building incredibly useful models through post-training and RL techniques.
You are a power user of AI models and eager to push the boundaries of what’s possible with reinforcement learning and alignment methods.
Compensation & Benefits
$180,000 - $600,000 USD
Base salary is just one part of our total rewards package at xAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short & long-term disability insurance, life insurance, and various other discounts and perks.
xAI is an equal opportunity employer. For details on data processing, view our
Recruitment Privacy Notice
.
Responsibilities
You will work on the most critical post-training and reinforcement learning challenges at any given time — including reward modeling, preference optimization (RLHF/DPO), and RL for improving reasoning, truthfulness, and real-world capabilities.
You will get clarity on your first project before an offer.