Nash
Staff Infrastructure and Performance Engineer
San Francisco · Staff+
Sponsorship not specifiedDetected 196 days ago
PostgreSQLAWSCI/CDIncident ResponseLogisticsShopifyLeadershipProblem Solving
About the role
- This is a senior, high-impact role focused on elastic capacity, high availability, cloud-native architectures, Postgres performance, and enterprise-grade CI/CD and multi-region deployments.
- You will set technical direction, define best practices, and deploy systems for the largest retailers in the world powering their business critical workflows.
Responsibilities
- Own infrastructure performance and reliability across Nash's production systems, with a focus on low latency, high throughput, and predictable behavior under load.
- Design, build, and optimize AWS-based infrastructure, leveraging managed services with a strong emphasis on ECS/Fargate.
- Lead Postgres performance engineering, including query optimization, indexing strategies, connection management, replication, cluster design, and failover.
- Design and evolve enterprise-grade CI/CD pipelines that support safe, repeatable, and fast deployments across environments and regions.
- Drive observability standards (metrics, logs, tracing, SLOs) and use data to proactively identify and eliminate performance bottlenecks.
- Partner with application engineers to influence system design decisions that impact scalability, latency, and reliability.
- Lead incident response and postmortems, focusing on root cause analysis, systemic fixes, and long-term resilience.
- As a Staff Infrastructure Performance & Engineer, you will own and evolve the performance, reliability, and scalability of Nash's core infrastructure.
- You'll work directly with the Engineering Leadership team, platform, and product engineering teams to design and operate low-latency, business-critical systems that power real-time logistics for some of the largest retailers in the world.
Requirements
- 6+ years of experience building and operating high-scale, production infrastructure for business-critical systems.
- Deep expertise in AWS, including networking, compute, storage, and managed services.
- Hands-on experience running production workloads on ECS/Fargate at scale.
- Proven experience designing and operating multi-region architectures with strict uptime and reliability requirements.
- Strong understanding of CI/CD for enterprise deployments, including rollout strategies, environment isolation, and rollback safety.
Compensation
- ✅ Competitive compensation and opportunity for equity
Company info
- Logistics is the substrate beneath every economy that has ever existed, and it remains the least intelligently coordinated activity in the modern world.
- Consumer expectations are converging on instantaneous, perfect, free.
- Networks are not.
- We call this gap the Logistics Singularity, and closing it is the work.
- Nash is the Autonomic Logistics OS.
- We unify decisioning, execution, and capacity into a single programmable system that runs the operation at equilibrium across orders, fleets, carriers, and customers.
- The world's largest retailers, grocers, and pharmacies (including Walmart, 7-Eleven, Woolworths, Coles, and Pet Circle) run their critical workflows on Nash.
- Founded by Mahmoud Ghulman and Aziz Alghunaim, and backed by Y Combinator, a16z, and other top investors.
- Headquartered in San Francisco.
This listing is sourced directly from Nash's careers page and normalized into a canonical job model.