AGI INC
Research Engineer - Evals
San Francisco Office
Sponsorship not specifiedDetected 56 days ago
ResearchLeadershipCollaboration
About the role
- Trustworthy, consumer-grade agents that redefine human-AI collaboration for millions.
- We're a stealth team of elite founders and AI researchers, with backgrounds spanning Stanford, OpenAI, and DeepMind.
- Models, agents, and product features all ship behind one question: did this actually get better?
Responsibilities
- BUILD THE FUTURE. ๐
- Build everyday AGI.
- Grounded in years of agent research, our AI is designed with trustworthiness and reliability as core pillars, not afterthoughts.
- Partnerships, by translating "did it get better" into language an OEM partner can hold us to
- The realities of shipping consumer agents to production partners
Company info
- Software shouldn't wait for commands; it should partner with you, amplifying what you can do every single day.
- WHY AGI, INC.
- We're industry leaders in mobile and computer-use agents, bringing these capabilities to consumer scale.
- We are supported by tier-1 investors who funded the first generation of AI giants; now they're backing us to build the next: everyday AGI. (Watch the demo https://drive.google.com/file/d/1ZydjdMeMh3x-QItUPQFbJUbhBW-4-XHa/view?usp=sharing)
- If you see possibility where others see limits, read on.
- You decide what "better" means.
- Without a strong evals function, the lab ships vibes.
- With one, every training run, every prompt change, every agent capability moves a number we trust - and the team makes decisions on real signal, not the loudest opinion in the room.
- You'll build the eval harness for AGI - across model capability, agentic behavior, on-device performance, and end-user experience.
Apply directly at AGI INC โCreate a free account for alerts like thisView AGI INC immigration profile
This listing is sourced directly from AGI INC's careers page and normalized into a canonical job model.