Machine Learning Engineer, Safety
Machine Learning Engineer, Safety | Stealth Mode Frontier AI Lab | Bay AreaI'm working with a well-funded, early-stage stealth AI lab building genuinely frontier systems - and they're hiring a Machine Learning Engineer focused on safety.The mission: make advanced AI systems reliable, controllable, and aligned as their capabilities grow. This is hands-on, unsolved-problem work at the edge of what's possible.What you'd work onEvaluation and oversight systems for advanced reasoning and agentic behaviourRed-teaming and adversarial testing - turning findings into real model and training improvementsSafety-focused post-training, reward modelling, and guardrailsIdentifying and mitigating failure modes in complex, multi-step reasoningYou might be a fit if you haveStrong ML engineering skills and hands-on experience with LLMs / foundation modelsWork in one or more of: post-training (SFT/RL/RLHF), evals, red-teaming, alignment, or safety infrastructureA bias toward shipping and owning problems end-to-end in an ambiguous environmentReal interest in the hard problems of frontier AI safetyDetailsBay Area, hybridSmall, senior, talent-dense team - real ownership from day oneHighly competitive compensation