Skip to main content
Nubank

Lead Systems Engineer (Kafka)

RemoteCanada only
Published
Role
DevOps
Experience
Lead
Employment
Full-time
Company size
Enterprise
$210k–$252k/yr
Check eligibility

Open to CA only. Set where you work from to check your eligibility.

No BS summary

Lead software/systems engineer in Canada for large-scale distributed infrastructure. Must have Kubernetes, AWS, networking, production reliability, observability, and troubleshooting experience. Kafka is desirable but not required.

Core skills

KubernetesAWSKafka

Optional skills

Apache Kafkamessaging technologiesstreaming technologies

What you'll do

  • Operate and improve large-scale messaging and platform infrastructure based on Kafka used by critical systems across Nubank
  • Contribute to the reliability, scalability, and performance of asynchronous communication platforms
  • Help design and implement solutions for high-throughput, low-latency, and fault-tolerant systems
  • Improve observability, automation, and operational excellence across the platform
  • Support incident analysis, troubleshooting, and root cause remediation in production environments
  • Optimize infrastructure usage and help drive efficiency and cost awareness across AWS-based environments
  • Work on platform capabilities that enable safe growth in message volume, topic count, and cluster footprint
  • Partner with other engineers and teams to evolve platform standards, tooling, and best practices
  • Contribute to architectural discussions involving messaging, traffic patterns, service communication, and platform reliability

What they require

  • Experienced software engineer to help evolve and operate Nubank’s messaging platform and the infrastructure that supports asynchronous communication at scale
  • At the Lead level, able to independently own important technical problems, improve reliability and operability, and drive engineering decisions in partnership with the team
  • Strong knowledge of distributed systems infrastructure, especially Kubernetes, networking, and AWS, is essential
  • Strong software engineering fundamentals and experience working with distributed systems in production
  • Solid experience with Kubernetes, networking, and AWS in large-scale or business-critical environments
  • Experience operating infrastructure-heavy platforms with high reliability and availability requirements
  • Ability to troubleshoot complex production issues across application, infrastructure, and network layers
  • Experience improving observability, automation, and operational tooling
  • Good understanding of scalability, resilience, performance, and failure isolation patterns
  • Ability to work autonomously on ambiguous technical problems and drive them to execution
  • Strong collaboration skills and ability to work across team boundaries
  • Preferred: Experience with platform engineering, SRE, or infrastructure-focused backend engineering
  • Preferred: Experience with high-throughput event-driven architectures
  • Preferred: Experience balancing reliability, performance, and cost in production systems

Benefits

  • Total compensation includes base salary, RSUs and benefits
  • Health Insurance
  • Life Insurance
  • Pension Plan
  • Extended maternity and paternity leaves
  • Nucleo - learning platform of courses
  • NuLanguage - language learning program
  • NuCare - mental health and wellness assistance program
  • Vacations
Financial ServicesEnterprisenubank.com.br

Details

Apply routeDom
$210k–$252k/yr