Anthropic
Machine Learning Infrastructure Engineer, Safeguards Research
San Francisco, CA | New York City, NY
H1B sponsorship available$350k-$500kDetected 13 hours ago
PythonDistributed SystemsMachine LearningData EngineeringNLPDetection EngineeringLogisticsResearchCommunication
About the role
- A growing part of that work depends on lightweight detection methods trained on model internals, which let us identify harmful behavior cheaply and at scale.
- This work feeds directly into Anthropic's Responsible Scaling Policy commitments.
- This is the tooling our researchers rely on to run experiments, train detection methods, and select detections for launch.
Responsibilities
- Own the training, evaluation, and scoring workflows researchers use, with a focus on cutting the time between an idea and a result
- Design tooling and interfaces, including libraries and command line tools, that researchers can use directly without needing to understand the systems underneath
- Build correctness and sanity checking into the stack, so results stay trustworthy as models and workloads evolve
- Partner closely with researchers and engineers across Safeguards to understand their workflows, anticipate how their needs will change, and design for that ahead of time
- Experience building and operating data-intensive or distributed systems in production
- Experience building tooling or infrastructure that other engineers or researchers use as a dependency
Requirements
- Strong software engineering fundamentals and hands-on coding ability, with proficiency in Python
- Comfort working across the research-to-deployment pipeline, from exploratory experiments to production systems
- Ability to debug performance and correctness problems across an unfamiliar stack
- Years of experience required will correlate with the internal job level requirements for the position
- Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience
- Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position
Nice to have
- Familiarity with language modeling and transformers, including working with model internals
- Experience with probes, interpretability, or classifier development
- Interest in the misuse risks of AI systems and a desire to work on mitigating them
Skills
- Take the highest-value research workflows from experiments to reliable, production-grade jobs
- Improve the throughput, cost, and reliability of large-scale inference and scoring workloads
Compensation
- $350,000 - $500,000 USD
Benefits
- Build and scale the infrastructure and data pipelines behind Safeguards machine learning research
- Minimum education: Bachelor's degree or an equivalent combination of education, training, and/or experience
Visa & Work Authorization
- However, we aren't able to successfully sponsor visas for every role and every candidate.
- But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.
- We do sponsor visas!
Apply directly at Anthropic →Create a free account for alerts like thisView Anthropic immigration profile
This listing is sourced directly from Anthropic's careers page and normalized into a canonical job model.