Skip to main content
Pinterest
Pinterest

Senior Site Reliability Engineer

RemoteUnited States only
Published
Role
SRE
Experience
Senior
$139.8k–$287.7k/yr
Check eligibility

Open to US only. Set where you work from to check your eligibility.

No BS summary

Senior SRE/DevOps engineer with 4+ years running production AWS and Kubernetes/EKS platforms. Must know ArgoCD GitOps, Helm, Terraform/Terragrunt, CI/CD, observability, Linux/containers/IAM/networking, and scripting in Bash or Python. US-based applicants only.

Core skills

AWSKubernetesArgoCD

Required skills

HelmTerraformTerragruntBash/PythonGitHub ActionsLinuxIAM

What you'll do

  • Ensure reliability, availability, and performance of production infrastructure and platform services
  • Operate and scale Kubernetes platforms, including governance and support for multi-tenant workloads
  • Manage GitOps-based deployment workflows using ArgoCD and Helm
  • Drive infrastructure provisioning and change management through Terraform/Terragrunt
  • Build and support CI/CD automation and deployment workflows using GitHub Actions
  • Lead incident response efforts, root cause analysis, and post-incident improvement initiatives
  • Reduce operational toil through scripting, tooling, and process automation
  • Advance observability practices across logs, metrics, traces, dashboards, and alerting
  • Support secure secrets integration, IAM-aware operations, and platform guardrails
  • Partner with application, security, and platform teams to improve reliability and delivery outcomes

What they require

  • 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Infrastructure
  • Strong hands-on experience operating AWS in production environments
  • Deep expertise in Kubernetes, including cluster operations, troubleshooting, workload reliability, and platform administration
  • Proven experience with Kubernetes multi-tenancy, including namespaces, RBAC, quotas, policies, and tenant isolation patterns
  • Experience implementing and operating ArgoCD within a GitOps delivery model
  • Strong hands-on experience with Helm
  • Strong experience with Terraform/Terragrunt for infrastructure provisioning and environment management
  • Solid scripting and automation skills using Bash and/or Python
  • Experience building, maintaining, or supporting CI/CD pipelines, ideally using GitHub Actions
  • Strong troubleshooting skills across Linux, containers, IAM, networking, and distributed systems
  • Experience with monitoring, alerting, and observability in production environments
  • Demonstrated ownership mindset with experience handling incidents, resolving production issues, and driving follow-through after outages
  • Strong collaboration and communication skills, with the ability to work effectively across engineering, security, and platform teams
  • Bachelor’s degree in computer science, engineering, a related field or equivalent experience
  • Demonstrated ability to use AI to improve speed and quality in day-to-day workflow for relevant outputs
  • Strong track record of critical evaluation and verification of AI-assisted work, such as testing, source-checking, data validation, and peer review
  • High integrity and ownership: protect sensitive data, avoid over-reliance on AI, and remain accountable for final decisions and deliverables

Benefits

  • Eligible for equity
  • Benefits available for this position are referenced in the posting

American photo sharing and publishing website

🇺🇸 United StatesSocial MediaEnterprisepinterest.com/

What people say about this company

3.8/ 5

Details

Apply routeDom
$139.8k–$287.7k/yr