Posted last month
Member of Technical Staff - Post-Training and RL at xAI in Palo Alto, CA. Work on critical post‑training and reinforcement learning challenges such as reward modeling and RL for reasoning.