Skip to main content
Serko

Principal Engineer - AI Platform & Operations

RemoteUnited States only
Published
Role
AI / ML
Experience
Principal
Employment
Full-time
$168k–$230k/yr
Check eligibility

Open to US only. Set where you work from to check your eligibility.

No BS summary

Principal-level engineer for ML infrastructure and AI platforms, with 12+ years engineering experience and 4+ years at Principal or Distinguished level. Must know model serving, Kubernetes, cloud ecosystems, MLOps tools, Python, Docker/Helm/containerization, and LLM inference at scale. Remote role based in Washington or California State.

Core skills

KubernetesLLM inferenceMLOps

Required skills

Triton/vLLM/Ray ServeAWS/GCP/AzureMLflow/Weights & Biases/KubeflowPythonDockerHelm

What you'll do

  • Define the long-term technical roadmap for the AI platform, covering model serving, feature stores, experiment tracking, and CI/CD for ML.
  • Establish engineering benchmarks for deployment, versioning, A/B testing, and automated rollbacks.
  • Lead strategies for GPU/compute efficiency and cost optimization while managing LLM inference complexities including quantization, batching, and latency at scale.
  • Design sophisticated monitoring and alerting systems tailored for AI workloads in production.
  • Drive platform stability and partner with application teams to ensure infrastructure meets evolving product needs.
  • Mentor Senior engineers, lead architecture reviews, and evaluate the next generation of cloud services and ML frameworks.
  • Build robust, self-serve internal developer platforms enabling engineers to deploy, monitor, and scale AI models efficiently and safely.

What they require

  • 12+ years of engineering experience, with at least 4 years at a Principal or Distinguished level.
  • Expert-level knowledge of model serving and deep experience with Kubernetes and cloud ecosystems.
  • Proven experience with tools like MLflow, Weights & Biases, or Kubeflow.
  • High proficiency in Python and a systems-thinking approach to Docker, Helm, and containerization.
  • Hands-on experience operating LLM inference at scale and a deep understanding of the trade-offs between throughput and latency.
  • A track record of building internal platforms that treat other engineers as the primary customer, drastically improving engineering velocity.
  • Thrives on solving complex, distributed systems problems.
  • Remote Role - can be based in either Washington or California State.

Benefits

  • A competitive base pay
  • Medical Benefits
  • Discretionary incentive plan based on individual and company performance
  • Access to a learning & development platform and opportunity for you to own your career pathways
  • Flexible work policy
  • Great tools and support to enable you to perform at the highest level of your abilities

Serko is a tech platform in global business travel & expense technology, offering a business travel marketplace.

TravelTech
$168k–$230k/yr