Skip to main content
Juniper Square

Senior Staff Software Engineer (AI)

RemoteUnited States, Canada only
Published
Role
Backend
Experience
Staff
Employment
Full-time
Company size
Enterprise
$265k+/yr
Check eligibility

Open to US, CA only. Set where you work from to check your eligibility.

No BS summary

Senior Staff Software Engineer with 15+ years of experience in distributed systems and platform engineering. Must have proven experience architecting large-scale, high-availability multi-tenant distributed systems and expertise in modern cloud-native stacks. Requires strong Python, Go, or Java skills and experience with infrastructure as code and event-driven architectures.

Core skills

distributed systemsplatform engineeringAI

Required skills

Kubernetescontainersservice meshserverlessPythonGoJavaTerraformPulumiCI/CDKafkaAWSAzureGCP

Optional skills

SaaS platform upmarketprivate marketsfund accountingfintech systemsAI/LLM workloadsinference servingagent runtimesGPU scheduling

What you'll do

  • Define and own the end-to-end systems architecture strategy across services, infrastructure, and developer platform
  • Designing frameworks and design patterns that accelerate the work of the rest of engineering org
  • Design scalable, fault-tolerant distributed systems spanning synchronous APIs, asynchronous workflows, and event-driven architectures
  • Establish standards for service design, API contracts, multi-tenancy, and inter-service communication
  • Lead architecture reviews and technical decision-making for high-stakes, cross-cutting initiatives
  • Drive adoption of modern architectures (event-driven systems, platform engineering, infrastructure as code)
  • Design and prototype critical platform and infrastructure components
  • Write production-quality code for complex or high-impact areas
  • Review service designs, infrastructure changes, and performance-critical code paths
  • Troubleshoot performance, scalability, and reliability issues across services, queues, caches, and databases
  • Optimize workloads for latency, throughput, concurrency, and cost
  • Design and architect a scalable cloud platform supporting compute, networking, storage, service orchestration, and deployment across all product lines
  • Build a paved-road developer platform — golden paths, internal tooling, CI/CD, and self-service infrastructure — that makes the right way the easy way
  • Create reusable platform primitives and internal APIs that product teams can compose rather than rebuild
  • Enable self-service infrastructure for engineering teams through standardized patterns, templates, and platform capabilities
  • Partner with AI workflow teams, product, and security teams to support production AI workloads, from inference pipelines to agent fleets
  • Design systems for safe agentic execution, including isolation boundaries, resource governance, auditability, and human-in-the-loop controls
  • Ensure reliability, scalability, and resilience of the platform, including high availability, monitoring, and disaster recovery readiness
  • Architect core systems to handle the data volumes and structural complexity of $100B+ AUM managers: thousands of entities, tens of thousands of LPs, and deep multi-tier fund structures
  • Define scalability targets and re-architect bottlenecked systems ahead of demand, ensuring performance holds at 10x current transaction and data volumes
  • Design enterprise integration surfaces — APIs, bulk data interfaces, and ERP/GL connectivity — that fit into the existing technology estates of large institutional GPs
  • Meet the security, compliance, and operational diligence expectations of the largest PE firms, such as, SSO/SCIM, granular entitlements, audit trails, and data residency
  • Serve as the senior technical voice in enterprise sales and onboarding conversations where architecture, scale, and resiliency are decision criteria
  • Lead incident response maturity: on-call practices, blameless postmortems, and systemic remediation
  • Drive capacity planning, load testing, and chaos/resilience engineering practices
  • Optimize cloud spend through architecture, workload placement, and continuous cost engineering
  • Mentor senior engineers across platform, infrastructure, and product teams
  • Partner with product, data, and business teams to align systems investments with company outcomes — including the GPX upmarket expansion
  • Translate business and product requirements into scalable, operable system designs
  • Influence roadmaps using platform, reliability, and cost considerations
  • Act as the executive technical authority for systems architecture and production engineering
  • Evangelize AI adoption, including AI-assisted engineering and operations
  • Help promote a culture of operational excellence and outcome-driven innovation
  • Become a role model for engineering excellence for the rest of the org

What they require

  • Advanced degree in Computer Science, Engineering, or related field
  • 15+ years in distributed systems, platform engineering, or infrastructure roles
  • Proven experience architecting large-scale, high-availability multi-tenant distributed systems in production
  • Strong hands-on experience with modern cloud-native stacks (Kubernetes, containers, service mesh, serverless)
  • Deep expertise in distributed systems fundamentals: consistency models, consensus, partitioning, idempotency, backpressure, and failure modes
  • Advanced proficiency in Python, Go, Java, or similar; strong systems-level debugging skills
  • Expertise in infrastructure as code (Terraform, Pulumi, or similar) and modern CI/CD
  • Experience designing event-driven and streaming architectures (Kafka or similar)
  • Experience building developer platforms, internal tooling, or paved-road infrastructure at scale
  • Strong understanding of observability practices and SRE principles (SLOs, error budgets, incident management)
  • Hands-on experience with AWS, Azure, or GCP at production scale
  • Strong understanding of infrastructure security, multi-tenancy, and compliance best practices
  • Ability to operate at both executive and deeply technical levels
  • While this is a remote role, some travel is required — to in-person leadership meetups, team offsites, and to customer sites in support of enterprise engagements where deep technical presence matters.

Benefits

  • Health, dental, and vision care for you and your family
  • Life insurance
  • Mental wellness coverage
  • Fertility and growing family support
  • Flex Time Off in addition to company-paid holidays
  • Paid family leave, medical leave, and bereavement leave policies
  • Retirement saving plans
  • Allowance to customize your work and technology setup at home
  • Annual professional development stipend

Juniper Square is an Operations Partner for private markets, unifying technology, data, and fund administration services into a single platform for GPs.

🇺🇸 United StatesFintechEnterprise
$265k+/yr