Skip to main content
AlphaSense

Staff Site Reliability Engineer

RemoteIndia only
Published
Role
SRE
Experience
Staff
Company size
Enterprise
Salary not disclosed
Check eligibility

Open to IN only. Set where you work from to check your eligibility.

No BS summary

Highly experienced Staff Site Reliability Engineer needed to architect core reliability platforms, lead incident response, and drive SRE best practices. Requires 8+ years in SRE/DevOps, production SaaS experience, proficiency in Python/Go, cloud platforms (AWS/GCP/Azure), Kubernetes, networking, monitoring/alerting, and incident management.

Core skills

KubernetesSite Reliability EngineeringAIOps

Required skills

PythonGoAWSGCPAzureTCP/IPDNSHTTP/Sload balancingPrometheusGrafanaDatadogELKOTELcontinuous profiling

What you'll do

  • Architect Reliability Paved Paths: Build frameworks and self-service tooling that let teams own the reliability of their services in a “You Build It, You Run It” culture.
  • Lead AI-Driven Reliability: Drive our AIOps strategy — automating diagnostics, remediation, and proactive failure prevention.
  • Champion Reliability Culture: Embed SRE practices across engineering via design reviews, production readiness, and operational standards.
  • Incident Leadership: Act as Incident Commander during critical events, modeling operational excellence, and ensuring blameless postmortems lead to lasting improvements.
  • Advance Observability: Deliver end-to-end monitoring, tracing, and profiling (Prometheus, Grafana, OTEL, Continuous Profiling) to optimize performance proactively.
  • Mentor & Multiply: Elevate engineers across SRE and product teams through mentorship, technical guidance, and knowledge sharing.

What they require

  • 8+ years of experience in Site Reliability Engineering, DevOps, or a similar role, with at least 3+ of those years operating in a Senior+ SRE position
  • Strong background in running production SaaS systems at scale.
  • Proficiency in at least one programming/scripting language (Python, Go, or similar).
  • Hands-on expertise with cloud platforms (AWS, GCP, or Azure) and Kubernetes.
  • Deep understanding of networking fundamentals (TCP/IP, DNS, HTTP/S, load balancing).
  • Experience with monitoring & alerting (Prometheus, Grafana, Datadog, ELK).
  • Familiarity with advanced observability (OTEL, continuous profiling).
  • Proven incident management experience, including leading high-severity incidents and postmortems.
  • Strong troubleshooting skills across the full stack.
  • Excellent communication and collaboration skills.

Benefits

  • AlphaSense is an equal-opportunity employer. We are committed to a work environment that supports, inspires, and respects all individuals. All employees share in the responsibility for fulfilling AlphaSense’s commitment to equal employment opportunity.
  • AlphaSense does not discriminate against any employee or applicant on the basis of race, color, sex (including pregnancy), national origin, age, religion, marital status, sexual orientation, gender identity, gender expression, military or veteran status, disability, or any other non-merit factor. This policy applies to every aspect of employment at AlphaSense, including recruitment, hiring, training, advancement, and termination.
  • In addition, it is the policy of AlphaSense to provide reasonable accommodation to qualified employees who have protected disabilities to the extent required by applicable laws, regulations, and ordinances where a particular employee works.

AlphaSense delivers AI-driven market intelligence and search built on public and private content including equity research, company filings, event transcripts, expert calls, news, trade journals, and clients’ own research content.

🇺🇸 United StatesTechnologyEnterprisealphasense.net/
Also posted in 1 other channel
Salary not disclosed