Anthropic

Anthropic

Machine Learning Infrastructure Engineer, Safeguards Research

San Francisco, CA | New York City, NY

H1B sponsorship available$350k-$500kDetected 13 hours ago
PythonDistributed SystemsMachine LearningData EngineeringNLPDetection EngineeringLogisticsResearchCommunication

About the role

  • A growing part of that work depends on lightweight detection methods trained on model internals, which let us identify harmful behavior cheaply and at scale.
  • This work feeds directly into Anthropic's Responsible Scaling Policy commitments.
  • This is the tooling our researchers rely on to run experiments, train detection methods, and select detections for launch.

Responsibilities

  • Own the training, evaluation, and scoring workflows researchers use, with a focus on cutting the time between an idea and a result
  • Design tooling and interfaces, including libraries and command line tools, that researchers can use directly without needing to understand the systems underneath
  • Build correctness and sanity checking into the stack, so results stay trustworthy as models and workloads evolve
  • Partner closely with researchers and engineers across Safeguards to understand their workflows, anticipate how their needs will change, and design for that ahead of time
  • Experience building and operating data-intensive or distributed systems in production
  • Experience building tooling or infrastructure that other engineers or researchers use as a dependency

Requirements

  • Strong software engineering fundamentals and hands-on coding ability, with proficiency in Python
  • Comfort working across the research-to-deployment pipeline, from exploratory experiments to production systems
  • Ability to debug performance and correctness problems across an unfamiliar stack
  • Years of experience required will correlate with the internal job level requirements for the position
  • Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience
  • Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position

Nice to have

  • Familiarity with language modeling and transformers, including working with model internals
  • Experience with probes, interpretability, or classifier development
  • Interest in the misuse risks of AI systems and a desire to work on mitigating them

Skills

  • Take the highest-value research workflows from experiments to reliable, production-grade jobs
  • Improve the throughput, cost, and reliability of large-scale inference and scoring workloads

Compensation

  • $350,000 - $500,000 USD

Benefits

  • Build and scale the infrastructure and data pipelines behind Safeguards machine learning research
  • Minimum education: Bachelor's degree or an equivalent combination of education, training, and/or experience

Visa & Work Authorization

  • However, we aren't able to successfully sponsor visas for every role and every candidate.
  • But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.
  • We do sponsor visas!

This listing is sourced directly from Anthropic's careers page and normalized into a canonical job model.