Skip to main content
NeoHPC

Lead Infrastructure Engineer

RemoteUnited States only
Published
Role
DevOps
Experience
Lead
Employment
Full-time
Company size
Startup
Salary not disclosed
Check eligibility

Open to US only. Set where you work from to check your eligibility.

No BS summary

Senior infrastructure engineer with 9+ years' experience, deep Kubernetes and Linux expertise, remote, to own the Kubernetes platform for a GPU-focused neocloud provider.

Core skills

Kubernetes

Required skills

LinuxInfrastructure as Codecontainerized workloads

What you'll do

  • Design, build, and operate production-grade Kubernetes infrastructure.
  • Design virtualized Kubernetes clusters and tooling for provisioning, upgrades, scaling, security, observability, and troubleshooting.
  • Build Infrastructure as Code and automation for repeatable provisioning and operations.
  • Design and operate Kubernetes networking, ingress, DNS, service discovery, load balancing, and network policies.
  • Design and operate Kubernetes storage, persistent volumes, and CSI-based systems.
  • Diagnose complex infrastructure problems across multiple layers of the stack.
  • Build internal platform services that improve reliability or developer experience.
  • Establish infrastructure standards, operational practices, observability, and reliability mechanisms.
  • Make pragmatic trade-offs between speed, reliability, simplicity, and maintainability.
  • Provide technical leadership and shape the platform architecture.

What they require

  • 9+ years of overall professional experience in infrastructure engineering.
  • Deep understanding of Kubernetes internals and core components.
  • Strong production experience operating Kubernetes in production.
  • Strong Infrastructure as Code and automation experience.
  • Strong understanding of Kubernetes networking, ingress, DNS, service discovery, load balancing, and storage.
  • Strong Linux and systems troubleshooting skills.
  • Experience operating containerized workloads in production.
  • Strong ability to debug across Kubernetes, Linux, networking, storage, virtualization, and physical infrastructure.
  • Experience building reliable, observable, and operable infrastructure.
  • Ability to own infrastructure problems from architecture through production.
  • Strong technical leadership and pragmatic decision-making.
  • 9+ years of experience in infrastructure engineering.
  • Strong production experience designing and operating Kubernetes environments.

NeoHPC is a small, remote-first neocloud provider building a full-stack GPU-focused edge platform from bare metal to inference services.

CloudStartup
Also posted in 1 other channel
Salary not disclosed