Skip to main content
Paddle

Site Reliability Engineer

RemoteUnited Kingdom only
Published
Role
SRE
Employment
Full-time
Salary not disclosed
Check eligibility

Open to GB only. Set where you work from to check your eligibility.

No BS summary

Site Reliability Engineer for Paddle’s platform team, focused on infrastructure, reliability, automation, AWS production systems, observability, and developer experience. Needs software development background, production operations experience, AWS ecosystem experience, Linux/networking knowledge, distributed systems, and monitoring tools.

Core skills

AWSOpenTelemetryTerraform

Required skills

LinuxHoneycombGrafanaPingdomIncident.io

What you'll do

  • Develop and maintain tools to maximise engineering efficiency; such as but not limited to automating deployment infrastructure and database upgrades
  • Seek out processes that can be improved with automation and have internal Developer Experience as a main driver. Collaborate and enable engineers to do their jobs more efficiently, working with other engineers on a regular basis
  • Create, maintain and test our system disaster recovery process, including tooling to automate the process
  • You’ll be able to choose from a selection of AI tools to support day-to-day work (e.g. code generation, investigation, automation, and documentation), and we’ll back sensible experimentation with the right guardrails.
  • Handle production incidents, author blameless postmortems and enrich operational playbooks and runbooks
  • Monitoring, alerting, and SLO tracking; hands-on SRE work, not just DevOps-style monitoring
  • Run performance investigations (load testing, bottleneck analysis) and drive tuning across apps, data stores and AWS.
  • Own cost optimisation workstreams: right-sizing, autoscaling policies, workload scheduling, storage tiering, and identifying waste across ECS/Fargate, RDS/Aurora, SQS and observability.
  • Be an advocate of the GitOps methodology

What they require

  • A software development background, with experience shipping and operating production services, plus strong fundamentals in testing, code review, CI/CD, and debugging.
  • Have experience working across the AWS ecosystem, partnering closely with AWS Solution Architects and subject-matter experts to design, review, and operate production systems
  • A curiosity about AI and how it’s reshaping software development.
  • Collaborative, security-minded, and detail-oriented. We move quickly, so you’ll thrive if you enjoy a fast-paced environment and take pride in doing things the right way.
  • Knowledge of platform and ops concepts such as networking and Linux administration
  • Experience working with microservices and distributed systems at scale
  • Experience with monitoring tools: we use Opentelemetry, Honeycomb, Grafana, Pingdom and Incident.io

Benefits

  • You can work remotely, from one of our stylish hubs, or even a bit of both!
  • Unlimited holidays
  • 4 months paid family leave regardless of gender
  • Annual learning fund
  • Regular internal and external training
  • Constant exposure to new challenges

Paddle offers SaaS companies payment infrastructure as a Merchant of Record, handling payment fragmentation for customers.

FintechMid-size
Salary not disclosed