Jobs
No opportunities available
There are currently no job opportunities available. Please check back later.
Research Engineer, Interpretability
Research Engineer, Interpretability
AnthropicHybridSFAI/MLFull-time$315K-$560K/yrGrowth
About the role
Anthropic is seeking a Research Engineer for their Interpretability team in San Francisco, CA, to reverse-engineer how trained models work and make advanced AI systems safe. The role focuses on mechanistic interpretability, aiming to discover how neural network parameters map to meaningful algorithms. Candidates should have 5-10+ years of software building experience, proficiency in languages like Python, and experience in empirical AI research. This position involves implementing and analyzing research experiments, optimizing workflows, and building tools to improve model safety.
Similar jobs
Browse Jobs by Role
Chief of StaffProduct MarketingForward Deployed EngineerForward DeployedSoftware EngineerProduct ManagerData ScientistDesignSalesMarketingOperationsEngineering ManagerFinanceCustomer SuccessHR & People OpsHardware EngineerStrategyAccountingFounding EngineerGrowthDevOps & InfrastructureMachine Learning Engineer
