TensorWave

TensorWave

Infrastructure Engineer – DevOps, Kubernetes & Automation

Las Vegas, Nevada · Senior · Contract

Sponsorship not specifiedDetected 47 days ago
PythonGoNode.jsGitCloud PlatformsKubernetesAnsibleCI/CDLinuxPrometheusGrafanaDevOpsEmbedded SystemsTest AutomationDNSFirewallLoad Balancing

About the role

  • This role will work across Ansible, Kubernetes, Linux systems, Git-based workflows, CI/CD tooling, and internal platform services.
  • The ideal candidate has strong Linux fundamentals, practical automation experience, and a desire to grow into deeper Kubernetes, DevOps, and infrastructure engineering responsibilities.

Responsibilities

  • Support cluster lifecycle activities including node maintenance, configuration updates, upgrades, and validation.
  • Support platform services that run on Kubernetes where owned by the infrastructure team.
  • Write, update, and maintain Ansible roles and playbooks.
  • Validate idempotency, error handling, and safe rollback behavior where applicable.
  • Assist with inventory organization, group variables, host variables, and reusable role design.
  • Support Git-based workflows for infrastructure code.
  • Help maintain deployment scripts, validation tooling, and operational utilities.
  • Support internal platform tooling used by the DevOps and infrastructure teams.
  • Support Ubuntu-based infrastructure systems and GPU node operating environments.
  • Help improve operational runbooks for common Linux and infrastructure support tasks.

Requirements

  • Linux system administration experience.
  • Experience troubleshooting services using logs, systemd, command-line tools, and metrics.
  • Ability to read and modify YAML, shell scripts, and infrastructure configuration files.

Nice to have

  • Experience with Ubuntu server environments.
  • Experience with RKE2, Rancher, Cilium, or similar Kubernetes platforms.
  • Experience with Prometheus, Grafana, Loki, or other observability tools.
  • Experience with MAAS, PXE, bare metal provisioning, or data center infrastructure.
  • Experience supporting GPU, AI, HPC, or large-scale compute environments.
  • Familiarity with Python or Go for operational tooling.
  • Experience working in production infrastructure environments with change control or staged rollout practices.
  • Life and Voluntary Supplemental Insurance Options

Benefits

  • 100% paid Medical, Dental, and Vision insurance for Employees
  • Company Health Savings Account Contributions
  • 100% paid Short Term and Long Term Disability Insurance for Employees
  • Various Supplementary Health Benefits, such as discounted Virtual Healthcare Appointments and Serious Illness Support
  • Flexible Spending Account
  • Parental Leave
  • Other In-Office Perks
  • Help investigate Kubernetes issues related to pods, services, networking, storage, ingress, certificates, and node health.

Company info

  • deliver seamless, secure, reliable, and resilient AI compute at scale.
  • We've built a versatile cloud platform that eliminates infrastructure barriers, empowering builders to focus on innovation instead of fighting their stack.
  • Because breakthrough AI should move at the speed of ideas, not infrastructure.

Equal opportunity

  • TensorWave is an Equal Opportunity Employer.
  • We celebrate diversity and are committed to creating an inclusive environment for all employees.
  • We do not discriminate on the basis of any protected status under applicable law.
  • Reasonable Accommodations

This listing is sourced directly from TensorWave's careers page and normalized into a canonical job model.