Skip to main content
RxSense

Principal Platform Engineer

RemoteUnited States only
Published
Role
DevOps
Experience
Principal
$190k–$225k/yr
Check eligibility

Open to US only. Set where you work from to check your eligibility.

No BS summary

Principal/architect-level platform engineer with 8+ years in platform, infrastructure, or SRE. Must be US-remote and work Eastern US business hours, with deep AWS, Kubernetes, CI/CD, IaC, observability, and hands-on coding in Python, Go, or TypeScript. Healthcare/regulatory experience is a plus.

Core skills

AWSKubernetesCI/CD

Required skills

IAMGitHubGitHub ActionsPython/Go/TypeScriptTerraform/CloudFormation/PulumiPrometheus/New Relic/Grafana/Datadog/OpenTelemetry

Optional skills

LLMGPU provisioningmodel servingagent runtime infrastructureKubernetes operatorscustom controllersKafkagRPC

What you'll do

  • Design, build, and operate the core cloud infrastructure that supports RxSense engineering, including compute, networking, identity and access management, and container orchestration on AWS.
  • Own the CI/CD platform used across engineering teams, including build pipelines, artifact management, environment promotion, progressive rollout, and rollback.
  • Build and maintain the observability stack across the organization, including logging, metrics, distributed tracing, alerting, and the SLIs and SLOs that define reliable service.
  • Drive infrastructure as code practices across the organization and manage the Kubernetes clusters, custom controllers, and operators that back production workloads.
  • Partner with security and compliance to manage secrets, network policy, and access controls appropriate to a regulated healthcare and pharmacy benefits environment.
  • Design self-service developer tooling and platform APIs and set the abstractions that let product and AI engineering teams’ provision, deploy, and operate services without handholding.
  • Own incident response tooling, on-call practices, and reliability engineering standards, and lead or support major incident response across the platform.
  • Partner with engineering leadership and finance on infrastructure cost visibility, capacity planning, and cost optimization.
  • Write documentation, runbooks, and clear interfaces so the platform is adoptable by other engineering teams without handholding.
  • Participate in code review and promote collaboration and best practices, including simplicity, automation, sound design patterns, test coverage, and reusability.
  • Work normal day business hours in the Eastern US time zone

What they require

  • BS (or higher, e.g., MS or Ph.D.) in Computer Science or a related technical field involving coding, or equivalent technical experience.
  • 8+ years of platform, infrastructure, or site reliability engineering experience, with demonstrated Staff, Principal, or Architect-level scope, such as owning platform architecture end to end or being accountable for critical production systems.
  • Deep, hands-on experience designing and operating CI/CD pipelines for high-velocity engineering organizations, including artifact management, environment promotion, and progressive rollout.
  • Strong AWS background, comfortable down to the IAM, networking, and container orchestration layers, including production Kubernetes experience.
  • Experience with GitHub and GitHub actions.
  • Proven track record building internal developer platforms or tools that other engineering teams adopted by choice, not by mandate.
  • Hands-on coding fluency in Python, Go, or TypeScript.
  • Comfortable operating in a polyglot environment.
  • The RxSense engineering stack spans Python, .NET, and TypeScript, and you will support services across all three.
  • Practical experience with infrastructure as code (such as Terraform, CloudFormation, or Pulumi) and modern observability tooling (such as Prometheus, New Relic, Grafana, Datadog, or Open Telemetry).
  • Comfortable owning the cost and reliability conversation with both engineering leadership and finance partners.
  • Strong written communication and a bias toward documentation, runbooks, and clear interfaces.
  • Proven analytical thinking and problem-solving skills.
  • Excellent communication skills, both verbal and written.
  • Preferred: Experience supporting infrastructure for AI or LLM-powered workloads, including GPU provisioning, model serving, or agent runtime infrastructure.
  • Preferred: Background in healthcare, PBM, pharmacy, or another regulated data environment.
  • Preferred: FinOps experience, particularly attributing infrastructure spend to features, teams, or business units.
  • Preferred: Experience with Agile development methodologies, preferably both Scrum and Kanban.

Healthcare Technology

Healthcare

Details

Apply routeGreenhouse
$190k–$225k/yr