Skip to main content
Teleport

Senior Site Reliability Engineer

RemoteUnited States only
Published
Role
SRE
Experience
Senior
Employment
Full-time
$222k–$326k/yr
Check eligibility

Open to US only. Set where you work from to check your eligibility.

No BS summary

Senior SRE with 5+ years of experience in Linux, networking, containers, and troubleshooting. Must have strong Go and Kubernetes development experience, and experience with systems observability tools. AWS Cloud experience is preferred. Willingness to complete a coding challenge in Go is required.

Core skills

GoKubernetesObservability

Required skills

Linux systemsnetworkingcontainerstroubleshootingPrometheusGrafanaLoki

Optional skills

AWS CloudGCP

Required languages

English

What you'll do

  • Re-engineer the core teleport product to scale globally and optimize routing latency for teams distributed around the world
  • Re-write portions of the core Teleport product to enable our goals for the cloud product
  • Build out our monitoring and observability stack to alert us to production issues and minimize false positives so we can all get a good sleep at night
  • Work on automation to tackle and eliminate the highest toil activities
  • Execute on traditional operation challenges, such as patching, scaling, backup and restore, disaster recovery, and more
  • Investigate the outages and incidents our customers experience with our product
  • Participate in the on-call rotation to ensure 24/7/365 system uptime.

What they require

  • Willingness to collaboratively work with Teleports’ engineers on coding challenge in Go https://github.com/gravitational/careers/blob/main/challenges/sre/challenge.md as part of the interview process.
  • 5+ years of progressive experience in Software Engineering and/or SRE/DevOps roles.
  • Strong experience in Linux systems, networking, containers, and troubleshooting.
  • Have solid Go and Kubernetes development experience.
  • Strong experience developing scripts, automation, or lightweight programs, submitting patches to the product codebase, or building tooling that incorporates AI agents into operational workflows.
  • AWS Cloud experience is preferred, GCP experience is acceptable.
  • Systems Observability tools: Prometheus, Grafana, Loki etc.
  • Operate and support the observability platform to maintain visibility and reliability.
  • Experience operate in a team where sound security choices are critical, and where reasoning about correctness and system invariants (e.g. formal or property-based methods) is valued.
  • Intellectual curiosity and a willingness to master new technologies.
  • Transparency, honesty, and a no-ego mindset.
  • Excellent communication skills.

Benefits

  • Extensive health coverage
  • Annual expense budget
  • Rest and recovery policies that maximize your ability to recharge
  • Investment in your future with retirement savings plans
  • Professional development opportunities

Product IT company developing solutions at the intersection of fintech, digital assets, and distributed systems.

$222k–$326k/yr