Anthropic Fellows Program, AI Safety & Security AnthropicRemote-Friendly (US) · London, UK +1 Red Team & SafeguardsRemoteNew 2w ago Not published View
Software Engineer, Safeguards Infrastructure AnthropicLondon, UK Red Team & SafeguardsNew 2w ago Not published View
Research Scientist, Frontier Risk Evaluations Scale AISan Francisco, CA · New York, NY Evals & BenchmarksNew 3w ago $216k – $270k View
Research Scientist, Agent Robustness Scale AISan Francisco, CA · New York, NY Red Team & SafeguardsNew 3w ago $216k – $270k View
Research Scientist, Safety Post Training Scale AISan Francisco, CA · New York, NY Post-Training & RLNew 3w ago $216k – $270k View
ML Systems Research Engineer, Agent Post-training Scale AISan Francisco, CA · New York, NY Post-Training & RLNew 3w ago Not published View