AGI INC

AGI INC

Research Engineer - Evals

San Francisco Office

Sponsorship not specifiedDetected 56 days ago
ResearchLeadershipCollaboration

About the role

  • Trustworthy, consumer-grade agents that redefine human-AI collaboration for millions.
  • We're a stealth team of elite founders and AI researchers, with backgrounds spanning Stanford, OpenAI, and DeepMind.
  • Models, agents, and product features all ship behind one question: did this actually get better?

Responsibilities

  • BUILD THE FUTURE. ๐Ÿš€
  • Build everyday AGI.
  • Grounded in years of agent research, our AI is designed with trustworthiness and reliability as core pillars, not afterthoughts.
  • Partnerships, by translating "did it get better" into language an OEM partner can hold us to
  • The realities of shipping consumer agents to production partners

Company info

  • Software shouldn't wait for commands; it should partner with you, amplifying what you can do every single day.
  • WHY AGI, INC.
  • We're industry leaders in mobile and computer-use agents, bringing these capabilities to consumer scale.
  • We are supported by tier-1 investors who funded the first generation of AI giants; now they're backing us to build the next: everyday AGI. (Watch the demo https://drive.google.com/file/d/1ZydjdMeMh3x-QItUPQFbJUbhBW-4-XHa/view?usp=sharing)
  • If you see possibility where others see limits, read on.
  • You decide what "better" means.
  • Without a strong evals function, the lab ships vibes.
  • With one, every training run, every prompt change, every agent capability moves a number we trust - and the team makes decisions on real signal, not the loudest opinion in the room.
  • You'll build the eval harness for AGI - across model capability, agentic behavior, on-device performance, and end-user experience.

This listing is sourced directly from AGI INC's careers page and normalized into a canonical job model.