Skip to main content
Palo Alto Networks

Sr Staff Site Reliability Engineer

RemoteBulgaria only· except United States
Published
Role
SRE
Experience
Senior
Employment
Full-time
Company size
Enterprise
Salary not disclosed
Check eligibility

Open to BG only · except US. Set where you work from to check your eligibility.

No BS summary

Senior SRE with 5+ years experience to operate large-scale multi‑cloud (GCP/AWS/Azure) production systems. Strong Kubernetes, Terraform and Python automation skills; experience with Prometheus/Grafana and incident response. Remote role based in/near CET timezones; employer does NOT sponsor visas.

Core skills

KubernetesTerraformPython

Required skills

GCPAWSAzurePrometheusGrafanaCI/CDGitOpsPagerDuty

Optional skills

GitLab CIGitHub ActionsJenkinsFlux

What you'll do

  • Own and operate large-scale, global production environments across multiple cloud providers (GCP, AWS, Azure)
  • Actively monitor, investigate, and resolve incidents triggered by automated alerting systems (PagerDuty / Incident Response)
  • Drive end-to-end troubleshooting across complex, distributed systems with high context switching
  • Design, deploy, and improve monitoring and observability systems (e.g., Prometheus, Grafana)
  • Collaborate closely with internal teams (CX, CS, Engineering) to ensure system reliability and performance
  • Work hands-on with modern DevOps and infrastructure tools including Kubernetes, Terraform, CI/CD pipelines, and GitOps workflows
  • Develop and maintain automation and tooling (primarily in Python)
  • Gain deep understanding of system architecture and interconnected services
  • Contribute to a culture of operational excellence in a high-scale, high-availability environment
  • Champion asynchronous communication, documentation, and tooling standards
  • On call responsibilities: Daytime hours (12:00–20:00 CET/CEST, based on candidate location and team coverage needs); Occasional weekends and holidays (rotation-based)

What they require

  • 5+ years of experience in SRE roles in production environments at scale
  • Strong hands-on experience with Kubernetes and Terraform
  • Strong hands-on experience with at least one major cloud platform (GCP or AWS required)
  • Experience building and configuring monitoring systems (e.g., Prometheus, Grafana)
  • Familiarity with CI/CD and GitOps tools (GitLab CI, GitHub Actions, Jenkins, Flux)
  • Proficiency in Python for scripting and automation
  • Proven success in a fully remote or distributed team environment
  • Strong troubleshooting and problem-solving skills with a passion for incident handling
  • Ability to work in fast-paced environments with high context switching
  • Highly responsive, proactive, and ownership-driven
  • Strong collaboration and communication skills

Benefits

  • Reasonable accommodations available for qualified individuals with a disability (contact accommodations@paloaltonetworks.com)
  • Equal opportunity employer; diversity and inclusion statements

At Palo Alto Networks®, we’re united by a shared mission—to protect our digital way of life.

🇺🇸 United StatesCybersecurityEnterprisepaloaltonetworks.com

Details

Visa sponsorshipNo
Salary not disclosed