Member of Technical Staff - Post-Training and RL na xAI

Presencial - Palo Alto, CA

Candidatar-se
Ver mais vagas na xAI

SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our small, highly motivated, engineering-focused team values hands‑on contributions, strong prioritization, communication skills, and leadership that emerges from initiative. The organization operates with a flat structure, encouraging all employees to contribute directly to the mission.

Salary

USD 180,000 - 600,000

Requirements

Skills

  • Experience with post-training, RLHF, or large-scale trained models (not required)
  • Experience with reinforcement learning and alignment methods
  • Power user of AI models
  • Belief in truth-seeking AI as a core challenge
  • Thrives in meritocratic environments and takes pride in work

Responsibilities

  • Work on critical post-training and reinforcement learning challenges including reward modeling, preference optimization (RLHF/DPO), and RL for improving reasoning, truthfulness, and real-world capabilities
  • Obtain clarity on first project before offer

Technologies

Post-trainingReinforcement learningReward modelingPreference optimizationRLHFDPOAlignment methods

Compartilhar vaga

Descubra se seu currículo está pronto para esta vaga

Veja como nossa IA pode otimizar seu currículo e aumentar suas chances de conseguir esta posição.