Skip to main content
Red Hat

Customer Site Reliability Engineer - OpenShift Managed Cloud Services (Spoken Japanese, Kubernetes/AWS/Azure, Linux)

RemoteAustralia only
Published
Role
SRE
Employment
Full-time
Company size
Enterprise
Salary not disclosed
Check eligibility

Open to AU only. Set where you work from to check your eligibility.

No BS summary

Customer Site Reliability Engineer with advanced experience in OpenShift/Kubernetes, Linux, and public clouds (AWS, Azure, or GCP) is needed to ensure availability, reliability, and performance of critical services. The role requires strong troubleshooting skills, a customer-first mindset, and exceptional communication, with spoken Japanese being an advantage. Must be based in Australia.

Core skills

OpenShiftKubernetesLinux

Required skills

AWS/Azure/GCP

Optional skills

PrometheusAnsibleTerraformGo

Required languages

English Fluent

Optional languages

JapaneseChineseKoreanSpanish

What you'll do

  • Manage large-scale, distributed systems, focusing on minimizing downtime and improving system resilience.
  • Maintain customer trust and confidence by ensuring stability and functionality of services.
  • Drive continuous enhancement of processes, tools, and methodologies to support the evolving needs of the service.
  • Lead the development of code and automation scripts to optimize the scalability, reliability, and performance of services.
  • Lead and participate in high-priority customer escalations, adopting a customer-first mindset.
  • Coordinate and execute complex incident response procedures, ensuring timely resolution and thorough postmortems.
  • Collaborate with cross-functional teams to enhance system robustness.
  • Demonstrate a proactive mindset to help preempt escalations and ensure reliable operations.
  • Document resolutions, root causes, and best practices to enrich the knowledge base and promote self-service solutions.
  • Mentor and coach team members, fostering a culture of continuous learning, knowledge sharing and collaboration.
  • Participate in on-call rotation and provide leadership during critical incidents.
  • Collaborate on strategic AI and automation projects designed to increase the efficiency of fleet operations and troubleshooting, ultimately delivering a better product experience for customers.

What they require

  • Advanced Experience with OpenShift/Kubernetes container platform support or administration.
  • Proficient with container-based technologies on Linux.
  • Proficient in managing Linux-based systems in a public cloud such as AWS, Azure, or GCP.
  • Advanced experience with enterprise systems monitoring;
  • Advanced with enterprise configuration management such as Ansible, Terraform.
  • Software engineering experience using object-oriented languages;
  • Superior communications skills and experience working directly with and presenting to customers.
  • Ability to quickly learn new technologies and follow industry trends.
  • Demonstrated ability to quickly and accurately troubleshoot systems issues.
  • Solid understanding of standard TCP/IP networking and common protocols.
  • Fluent in English and any additional language like Japanese, Chinese, Korean, Spanish is an advantage.

Red Hat is the world’s leading provider of enterprise open source software solutions, using a community-powered approach to deliver high-performing Linux, cloud, container, and Kubernetes technologies.

🇺🇸 United StatesEnterprise SoftwareEnterpriseredhat.com/
Salary not disclosed