SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our small, highly motivated, engineering-focused team values hands‑on contributions, strong prioritization, communication skills, and leadership that emerges from initiative. The organization operates with a flat structure, encouraging all employees to contribute directly to the mission.
Member of Technical Staff - Post-Training and RL na xAI
Presencial - Palo Alto, CA
Ver mais vagas na xAISalary
USD 180,000 - 600,000
Requirements
Skills
- Experience with post-training, RLHF, or large-scale trained models (not required)
- Experience with reinforcement learning and alignment methods
- Power user of AI models
- Belief in truth-seeking AI as a core challenge
- Thrives in meritocratic environments and takes pride in work
Responsibilities
- Work on critical post-training and reinforcement learning challenges including reward modeling, preference optimization (RLHF/DPO), and RL for improving reasoning, truthfulness, and real-world capabilities
- Obtain clarity on first project before offer
Technologies
Post-trainingReinforcement learningReward modelingPreference optimizationRLHFDPOAlignment methods
Descubra se seu currículo está pronto para esta vaga
Veja como nossa IA pode otimizar seu currículo e aumentar suas chances de conseguir esta posição.