San FranciscoFull TimeEngineering
Remotely
machine learningsimulationpublicationsbenchmarksexperiments
Job Description
📋 Description Tackle long-horizon evaluation for autonomous agents Design simulation environments for safe, autonomous agents Build benchmarks and run rigorous experiments Write production-quality code and ship results Publish research and share findings with the team Wear multiple hats as founding team member 🎯 Requirements Strong engineering and research fundamentals; proficiency with AI tools Experience post-training frontier models Experience shipping reliable, production-quality code Track record of publications 🎁 Benefits Comprehensive health, dental, and vision insurance Free meals with the team Free Bay Club membership nearby Commuter benefits 401(k) Unlimited PTO