Skip to main content
AZX

ML Engineer (Senior/Staff)

RemoteUnited States only
Published
Role
Fullstack
Experience
Senior
Employment
Full-time
$140k–$220k/yr
Check eligibility

Open to US only. Set where you work from to check your eligibility.

No BS summary

We're looking for an ML Engineer to own the technical backbone of how AZX serves and evaluates models at scale.

Core skills

vLLM/SGLang/TensorRT-LLM

Required skills

KubernetesGPU schedulingautoscaling

Optional skills

PyTorch LightningJAXHuggingFaceONNXC++RustCUDASpark

Required languages

English unknown

What you'll do

  • Own architecture for inference serving and GPU scheduling — Kubernetes operators, autoscaling, and dynamic capacity across vLLM/SGLang deployments on cloud and customer-managed infrastructure.
  • Design and calibrate eval systems for model, prompt, and agent changes, including golden datasets, LLM-as-judge pipelines, and regression gates wired into CI.
  • Advise on cost-aware model routing and cascading decisions, balancing latency, cost, and quality across providers and model tiers.
  • Apply physics-informed ML and enterprise AI expertise to the hardest client and platform problems, drawing on the team's research depth.
  • Set technical standards for ML infrastructure and evaluation practice across the org, and mentor engineers working in this space.
  • Partner closely with the inference platform, gateway, and evals-focused engineers to keep architecture coherent as the platform grows.

What they require

  • 3+ years of experience with ML infrastructure and inference serving — vLLM, SGLang, TensorRT-LLM, or comparable systems — at production scale.
  • Strong background in evaluation and reliability engineering for ML/LLM systems, or the seniority to build this practice from scratch.
  • Solid Kubernetes experience, ideally including GPU-specific scheduling constraints (node pools, autoscaling under GPU bottlenecks).
  • A track record of technical leadership at a staff or senior level — setting direction, not just executing tickets.
  • Research fluency is a plus (PhD, publications, or equivalent depth) given the technical bar of our existing ML team, though this is an infrastructure-and-systems role first.

Benefits

  • Competitive early-stage startup compensation (based on capabilities, experience, and location)
  • Bonus eligibility
  • Health insurance with meaningful coverage for dependents
  • Flexible paid time off
  • Equity
  • Fully remote culture with a cluster of teammates in Seattle

AZX

Our mission is to accelerate positive impact in critical industries through AI transformation. We’re growing quickly and already work with category-leaders in real estate (CBRE), energy (LevelTen Energy), logistics (Flexe), and utilities (Puget Sound Energy). AZX is a public benefit corporation founded in 2024. We were profitable through bootstrapped consulting for the first year. In early 2026, we raised $6M to scale our operations and technology. We work on challenges in clean energy, decarbonization, climate risk, energy systems, and global economics. We’re building our company for long-term success and aim to create the ultimate place to work for those passionate about AI and making a positive impact.

AI/ML For Climate And SustainabilityStartup

Details

Visa sponsorshipNo
$140k–$220k/yr