CyberArk

CyberArk

Principal Cloud Infrastructure Engineer (Advanced Threat Protection)

Office - USA - CA - Headquarters, USA · Principal · Full-time

No sponsorshipDetected 25 days ago
PythonBashDistributed SystemsGitPostgreSQLMySQLRedisBigQueryAWSGCPCloud PlatformsTerraformAnsibleCI/CDLinuxDevOpsSite Reliability EngineeringMachine LearningLLMsCommunicationCollaborationProblem SolvingCritical Thinking

About the role

  • Palo Alto Networks is at the forefront of cloud-native infrastructure, where reliability, scale, and intelligent automation define the future of operations.
  • This isn't just about keeping the lights on.
  • If you're excited about applying AI to real-world infrastructure challenges - and you thrive in an environment where automation isn't just a nice-to-have but a core philosophy - this is your next career.

Responsibilities

  • Design, build, and operate cloud infrastructure that enables reliable, rapid deployment of microservices with resilient operations and effective monitoring
  • Build and integrate AI-powered tools (e.g., LLM-based agents, AIOps platforms) into SRE workflows for intelligent alerting, log analysis, and capacity planning
  • Develop self-healing systems that can automatically detect anomalies, diagnose issues, and take corrective action with minimal human intervention
  • Identify and drive opportunities to improve automation for code deployment, management, and observability of application services
  • Lead root cause analysis of critical business and production issues, building runbooks and automation to prevent recurrence
  • Represent SRE in design reviews and work cross-functionally with engineering teams on operational readiness
  • You'll build intelligent systems that predict incidents before they happen, automate root cause analysis, and continuously optimize our infrastructure.
  • You'll be a critical bridge between engineering and our Infrastructure Platform, combining deep SRE expertise with AI-driven automation to deliver unprecedented levels of reliability and operational efficiency.

Requirements

  • 7+ years of experience in DevOps, Site Reliability, or infrastructure engineering
  • Expertise in multi-cloud environments. strong hands-on experience with GCP, AWS, and familiarity with OCI (Oracle Cloud Infrastructure)
  • Experience designing and operating infrastructure across multiple cloud providers, including networking, identity management, and cross-cloud connectivity
  • Expertise in Infrastructure as Code with tools such as Terraform, Ansible
  • Strong proficiency in Python and shell scripting for automation
  • Strong experience with Linux and distributed systems handling high-volume transactions
  • Familiarity with CI/CD pipelines, GitLab, and Artifactory

Nice to have

  • Experience applying AI/ML to operational workflows (e.g., AIOps, intelligent alerting, automated remediation, or LLM-powered tooling) is a strong plus
  • Experience with cloud compliance frameworks (FedRAMP, IL5) and operating in regulated environments is a plus

Compensation

  • The compensation offered for this position will depend on qualifications, experience, and work location.
  • For candidates who receive an offer at the posted level, the starting base salary (for non-sales roles) or base salary + commission target (for sales/com-missioned roles) is expected to be the annual range listed below.
  • The offered compensation may also include restricted stock units and a bonus.
  • We're trailblazers that dream big, take risks, and challenge cybersecurity's status quo.

Benefits

  • A description of our employee benefits may be found here.

Company info

  • In order to be the cybersecurity partner of choice, we must trailblaze the path and shape the future of our industry. This is something our employees work at each day and is defined by our values: Disruption, Collaboration, Execution, Integrity, and Inclusion. We weave AI into the fabric of everything we do and use it to augment the impact every individual can have. If you are passionate about solving real-world problems and ideating beside the best and the brightest, we invite you to join us!
  • We believe collaboration thrives in person. That's why most of our teams work from the office full time, with flexibility when it's needed. This model supports real-time problem-solving, stronger relationships, and the kind of precision that drives great outcomes.
  • Leverage AI/ML to automate incident detection, root cause analysis, and remediation - reducing toil and accelerating mean time to resolution
  • Write automation code for provisioning and operating infrastructure at massive scale
  • Work with development teams to ensure applications are production-ready, scalable, and reliable from the ground up
  • Establish end-to-end monitoring and alerting on all critical components, incorporating AI-driven anomaly detection and predictive analytics
  • Participate in the on-call rotation supporting the platform and production applications
  • Mentor other SREs on best practices in infrastructure orchestration, production troubleshooting, and AI-augmented operations
  • At Palo Alto Networks®, we're united by a shared mission-to protect our digital way of life.
  • We thrive at the intersection of innovation and impact, solving real-world problems with cutting-edge technology and bold thinking.
  • Here, everyone has a voice, and every idea counts.
  • If you're ready to do the most meaningful work of your career alongside people who are just as passionate as you are, you're in the right place.
  • In order to be the cybersecurity partner of choice, we must trailblaze the path and shape the future of our industry.
  • This is something our employees work at each day and is defined by our values: Disruption, Collaboration, Execution, Integrity, and Inclusion.
  • We weave AI into the fabric of everything we do and use it to augment the impact every individual can have.
  • If you are passionate about solving real-world problems and ideating beside the best and the brightest, we invite you to join us!
  • We believe collaboration thrives in person.
  • That's why most of our teams work from the office full time, with flexibility when it's needed.
  • This model supports real-time problem-solving, stronger relationships, and the kind of precision that drives great outcomes.

Equal opportunity

  • If you require assistance or accommodation due to a disability or special need, please contact us at accommodations@paloaltonetworks.com.
  • equal opportunity employer.
  • All your information will be kept confidential according to EEO guidelines.

Visa & Work Authorization

  • Is role eligible for Immigration Sponsorship?
  • Please note that we will not sponsor applicants for work visas for this position.

This listing is sourced directly from CyberArk's careers page and normalized into a canonical job model.