Skip to main content
Nagarro

Senior Site Reliability Engineer (AWS Cloud)

RemoteRomania only
Published
Role
SRE
Experience
Lead
Employment
Full-time
Company size
Enterprise
Salary not disclosed
Check eligibility

Open to RO only. Set where you work from to check your eligibility.

No BS summary

Senior SRE with 8+ years of cloud experience, specifically AWS. Must have strong Kubernetes, Terraform, and CI/CD pipeline skills. Experience with incident management and SLOs is crucial. Ability to simplify complex systems and collaborate with product teams is key.

Core skills

AWSKubernetesTerraform

Required skills

Infrastructure as Code (IaC)CI/CD pipelinesSLIs/SLOsmonitoringobservabilityIncident ManagementRoot Cause Analysis (RCA)

What you'll do

  • Architect & Drive the reliability, scalability, and performance of our multi-cloud provisioning platform across all production stacks.
  • Architect and implement end-to-end automation pipelines to eliminate manual intervention, actively identifying and reducing technical toil.
  • Define, monitor, and improve critical system health indicators (SLIs/SLOs), including latency, throughput, error rates, and capacity usage, making data-driven architectural recommendations.
  • Lead Collaboration with product and cross-functional engineering teams to embed reliability and security considerations early into the software development lifecycle (SDLC).
  • Own Incident Response Management: Design robust detection mechanisms, triage critical incidents, lead deep Root Cause Analysis (RCA), and implement long-term preventative engineering solutions.
  • Simplify Complex Systems: Continually audit platform operations to identify bottlenecks, eliminate single points of failure, and reduce structural complexity.

What they require

  • 8+ years of experience working within the cloud environment, in roles such as SRE (Site Reliability Engineer) or Cloud Platform/Reliability Engineer
  • Strong experience in cloud development and multi-cloud environments, preferably with a strong exposure to AWS cloud
  • Knowledge of cloud architecture, scalability, and high-availability design
  • Hands-on experience with Kubernetes and container orchestration
  • Experience with Terraform and Infrastructure as Code (IaC)
  • Experience designing automation and CI/CD pipelines to reduce operational toil
  • Strong understanding of SRE principles, SLIs/SLOs, monitoring, and observability
  • Proven experience with Incident Management, Root Cause Analysis (RCA), and reliability engineering
  • Ability to identify performance bottlenecks, single points of failure, and architectural risks

global digital engineering with full-service offering

πŸ‡ΊπŸ‡Έ United StatesIT ServicesEnterprisenagarro.com/

What people say about this company

3.7/ 5

Salary not disclosed