Skip to main content
Mrsool

Site Reliability Engineer II

RemoteIndia only
Published
Role
SRE
Experience
Senior
Employment
Full-time
Salary not disclosed
Check eligibility

Open to IN only. Set where you work from to check your eligibility.

No BS summary

Experienced SRE with 5+ years in high-traffic, high-availability environments. Must know at least one programming language, cloud infrastructure, Kubernetes/Docker, automation/config management, monitoring, distributed systems and networks. Role is based in India.

Core skills

AWS/GCP/Azure/Kubernetes/DockerSite Reliability EngineeringCloud Infrastructure

Required skills

Python/Ruby/Java/GoChef/Ansible/Puppet/TerraformPrometheus/Grafana/Nagios

Optional skills

DatabasesBackend technologiesBackend frameworks

What you'll do

  • Collaborate with development teams to design and implement scalable Infrastructure.
  • Collaborate with development teams to design and implement automated deployment and testing pipelines.
  • Develop and maintain monitoring and alerting systems to proactively identify and address issues.
  • Troubleshoot and escalate production incidents to minimize downtime and improve system reliability.
  • Continuously improve our infrastructure and processes to optimize scalability and efficiency.
  • Participate and take ownership for on-call rotations as needed to ensure 24/7 support for our application.
  • Perform routine maintenance and upgrades as needed to keep our systems up to date.
  • Contribute to ongoing efforts to improve our security posture and compliance with industry standards.
  • Communicate complex technical concepts clearly and concisely to both technical and non-technical stakeholders in order to make the right decision.
  • Mentor and coach junior engineers, fostering their professional growth and enabling them to deliver high-quality work.
  • Stay up-to-date with the latest advancements and trends in site reliability engineering and share knowledge and insights with the team.
  • Identify opportunities for organizational enhancements and propose alternatives to optimize team structures and execution.

What they require

  • Bachelor’s degree in Computer Engineering, Computer Science, or related field.
  • 5+ years of experience in a similar role, preferably with experience in a high-traffic, high-availability environment.
  • Proficiency in at least one programming language (Python, Ruby, Java, Go, etc.).
  • Strong understanding of cloud infrastructure and related technologies (AWS, GCP, Azure, Kubernetes, Docker, etc.)
  • Excellent troubleshooting and problem-solving skills.
  • Experience with one or more automation and configuration management tools (Chef, Ansible, Puppet, Terraform, etc.).
  • Familiarity with monitoring and alerting tools (Prometheus, Grafana, Nagios, etc.)
  • Strong communication and interpersonal skills, enabling effective collaboration with cross-functional teams.
  • Ability to navigate ambiguity, set clear expectations, and thrive in a fast-paced, dynamic environment.
  • A strong grasp of computer science fundamentals when it comes to dealing with distributed systems and networks.
  • Preferred: Experience or familiarity with running, tuning, and optimizing databases and queries.
  • Preferred: Prior experience in a leadership or senior-level role within site reliability engineering.
  • Preferred: Familiarity with other backend technologies and frameworks.
  • Preferred: Experience in driving technical decisions and implementing positive organizational changes.

Benefits

  • Inclusive and Diverse Environment: We foster an inclusive and diverse workplace that values innovation and offers remote environments.
  • Competitive Compensation: Our compensation packages are highly competitive and include potential share options for certain roles.
  • Personal Growth and Development: We are committed to your personal and professional growth, providing regular training and an annual learning stipend to help you advance your career in a dynamic environment.
  • Autonomy and Mentorship: You'll enjoy a high degree of autonomy in your role, supported by mentorship and ambitious goals that pave the way for both your success and the company's growth.

Mrsool is one of the largest delivery platforms in the Middle East and North Africa (MENA) region, offering an “order anything from anywhere” on-demand delivery experience powered by dedicated couriers.

Delivery
Salary not disclosed