Cerebras Systems
AI Inference Core - Senior SW Engineer for Platform & DevOps
Sunnyvale, CA · Senior
Stay score
odds of building a lasting career here
Thin sponsorship signal and lottery-bound. A low-probability bet with your clock running. Prioritize cap-exempt roles and proven entry-level sponsors first.
Lottery odds assume a STEM candidate.
Personalize to your clock →Employer immigration record
from this employer's Department of Labor filings
Green-card filing pattern in this occupation
Files H-1B transfers
Sourced from Department of Labor LCA, PERM and prevailing-wage disclosure data. Employer matching is by name, so figures may be split across an employer's legal entities. Absence of a filing means none appears in our copy of the data, not that none exists.
Community outcomes
No reports yet — be the first to help the next applicant.
About the role
- You will work on CI/CD systems, Kubernetes, deployment automation, cloud and on-premises infrastructure, developer environments, artifact management, and observability.
- You will help make the systems engineers depend on reliable, scalable, and easy to operate.
- This is an engineering-focused infrastructure role rather than a primarily ticket-driven operations position.
Responsibilities
- Design, build, and maintain CI/CD systems supporting build, test, integration, qualification, and release workflows.
- Build and operate Kubernetes-based platforms and services used by engineering teams across Cerebras.
- Develop deployment systems, internal tools, and self-service workflows that make infrastructure changes repeatable, reviewable, and safe.
- Perform root-cause analysis and implement lasting fixes rather than relying on repeated manual intervention.
- Partner with software, IT, security, networking, release, and developer-productivity teams to deliver scalable infrastructure solutions.
Requirements
- 5+ years of professional experience in platform engineering, DevOps, infrastructure engineering, site reliability engineering, or software engineering.
- Experience deploying and operating services using Kubernetes and containerized environments.
- Experience with a major cloud platform, preferably AWS, and programmatic infrastructure provisioning.
- Strong understanding of Linux or Unix operating-system fundamentals.
- Proficiency in Python, Shell, or another language used to build infrastructure automation and operational tooling.
- Experience with monitoring, logging, alerting, dashboards, and incident investigation.
- Strong debugging and problem-solving skills, including the ability to investigate issues spanning applications, infrastructure, networking, and operating systems.
- Experience with infrastructure-as-code tools, specifically Terraform.
- Experience with Kubernetes controllers, operators, custom resources, Helm, Argo CD, or similar platform technologies.
- Hands-on experience building or maintaining CI/CD pipelines and automated software-delivery workflows.
- Understanding of networking concepts such as DNS, routing, load balancing, proxies, ports, TLS, and service connectivity.
- Experience managing artifact repositories, package registries, build caches, or software-distribution infrastructure.
- Familiarity with build systems, dependency management, and reproducible-build practices.
- Experience supporting hybrid environments spanning cloud infrastructure, on-premises systems, and specialized hardware.
Skills
- Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups.
Company info
- The Core Infrastructure team builds the software systems that power engineering workflows across Cerebras.
- Our infrastructure coordinates complex work across machines, clusters, development environments, and hardware systems.
- We build orchestration frameworks, execution engines, scheduling systems, test infrastructure, developer tools, and reusable software platforms that allow engineers to build, test, qualify, and deliver software reliably at scale.
- These systems are primarily built in Python, but the work goes far beyond scripting or automation.
- Our frameworks act as the control plane for distributed workflows, managing resources, execution state, concurrency, failures, retries, dependencies, and observability across large and complex environments.
- Build a breakthrough AI platform beyond the constraints of the GPU.
- Publish and open source their cutting-edge AI research.
- Work on one of the fastest AI supercomputers in the world.
- Enjoy job stability with startup vitality.
- Our simple, non-corporate work culture that respects individual beliefs.
This listing is sourced directly from Cerebras Systems's careers page and normalized into a canonical job model.