Перейти к основному содержимому
Eltropy

SRE and Devops Team Manager

УдалённоIndia только
Опубликовано
Роль
DevOps
Опыт
Мидл
Занятость
Полная занятость
Размер компании
Средняя
Зарплата не указана
Проверьте доступность

Доступно для: IN only. Укажите, откуда вы работаете, чтобы проверить доступность.

Коротко по делу

Hands-on SRE/DevOps manager to lead a distributed team (IST + US) for a FinTech SaaS platform. Needs 2–3 years managing SRE/DevOps engineers, strong multi-cloud and IaC (Terraform/Pulumi) experience, incident management and observability skills. 100% remote but hiring limited to India (coordinates across IST and US time zones).

Ключевые навыки

KubernetesSRE

Обязательные навыки

SRE/DevOps managementincident managementAWSGKEEKSmanaged databasesobject storagenetworkingIAMTerraformPulumiCI/CDGitHub ActionsJenkinsobservabilitymetricsloggingtracingalerting

Желательные навыки

PulumiFinTech experiencedatabase operationscost-efficiency designresilience designsecurity design

Чем предстоит заниматься

  • Lead, mentor, and grow a team of SRE/DevOps engineers.
  • Oversee the incident management process end to end — on-call rotations, escalation paths, incident command, postmortems, and follow-through on RCA and preventative actions.
  • Define and drive SRE principles including SLIs, SLOs, error budgets, capacity planning, and observability standards across services.
  • Partner with engineering leadership to identify and implement process improvements across operational priorities.
  • Coordinate a distributed team and support model across IST and US time zones, ensuring effective handoffs, coverage, and communication.
  • Interface with engineering leadership to align infrastructure investments with business and compliance goals.
  • Report on team health, reliability metrics, and operational performance.
  • Assist engineering leadership in managing the team, including department-wide initiatives and roadmap planning.

Что требуется

  • 2-3 years of experience as a team lead or manager in a DevOps/SRE environment.
  • Proven experience leading and coordinating distributed teams across time zones, specifically IST and US, including on-call coverage and handoffs.
  • Experience running an incident management process — on-call, escalation, incident command, severity frameworks, and driving blameless postmortems and RCA to completion.
  • Strong hands-on background across multi-cloud environments (e.g., AWS, Kubernetes - GKE & EKS, managed databases, object storage, networking, IAM).
  • Solid experience with Infrastructure-as-Code practices (Terraform, Pulumi, or similar) and CI/CD tooling (GitHub Actions, Jenkins, or similar).
  • Experience with modern observability practices and tooling (metrics, logging, tracing, alerting).

Преимущества

  • Equal opportunity employer statement

Fintech SaaS Firm

FintechСтартап
Зарплата не указана