Skip to main content
Alpaca
Alpaca

Senior Data Engineer

RemoteNorth AmericaLATAM+United States, Canada
Published
Role
Data Engineering
Experience
Senior
Company size
Mid-size
Salary not disclosed
Check eligibility

Open to Anywhere in North America & LATAM + US, CA. Set where you work from to check your eligibility.

No BS summary

Senior Data Engineer with 5+ years in data engineering and 2+ years operating low-latency platforms handling 100M+ events/day. Must be strong with Kubernetes, Terraform/Ansible, streaming/CDC, Iceberg, Python, SQL, and GCP/cloud data infrastructure. Remote role hiring in North America and LATAM.

Core skills

KubernetesApache IcebergKafka

Required skills

DockerHelmTerraformAnsibleArgoCDTrino/PrestoRedpandaDebeziumAirflowAirbytePythonSQLGCPGCSCloud BuildCloud SQLDataproc

Optional skills

CubedbtLookerHightouchOpenMetadataDatahubApache Ranger

What you'll do

  • Design, build, and evolve core data platform infrastructure, including distributed query engines, orchestration, warehousing, and cataloging.
  • Own lakehouse infrastructure as code and manage deployments through Terraform and Ansible on Kubernetes.
  • Build and maintain low-latency streaming and CDC ingestion pipelines, plus batch ingestion paths landing in Iceberg.
  • Develop and scale the BI landscape so downstream teams and agents get performant, self-serve access to lakehouse data.
  • Enforce platform reliability best practices, including monitoring and alerting, on-call rotations, incident response, maintenance windows, runbooks, and SLAs.
  • Partner with DevOps, Analytics Engineering, and other stakeholders to close infrastructure gaps and support new data requirements.

What they require

  • 5+ years of experience in Data Engineering, including 2+ years building and operating scalable, low-latency data platforms handling more than 100M events per day.
  • Strong hands-on experience running data infrastructure on Kubernetes, with cloud-native tooling like Docker and Helm.
  • Production experience with infrastructure as code using Terraform, Ansible, and ArgoCD or equivalents.
  • Deep knowledge of distributed systems, including storage, transactions, and query processing.
  • Hands-on experience operating open-source query engines like Trino or Presto.
  • Strong experience with object storage and open table formats, specifically Apache Iceberg.
  • Experience with streaming and CDC systems: Kafka, Redpanda, and Debezium.
  • Hands-on experience with orchestration frameworks and ELT tools.
  • Strong working knowledge of Python and SQL for building pipelines and platform tooling.
  • Experience with Google Cloud Platform and its data services, or related experience with other cloud services.
  • Ability to thrive in a fast-paced startup environment and adapt infrastructure to rapidly changing needs.
  • Preferred: Experience with semantic/metrics layers.
  • Preferred: Familiarity with transformation frameworks.
  • Preferred: Familiarity with reverse ETL tooling.
  • Preferred: Familiarity with data catalog and lineage tooling.
  • Preferred: Experience with data access control and governance frameworks.

Benefits

  • Competitive salary and stock options.
  • Health benefits.
  • New hire home-office setup: one-time USD $500.
  • Monthly stipend: USD $150 per month via a Brex Card.

Alpaca is a modern platform for trading. Alpaca's API is the interface for your trading algorithms, bots, or applications to communicate with Alpaca''s brokerage and other services.

🇺🇸 United StatesFinanceMid-sizealpaca.markets

What people say about this company

3.6/ 5

Details

Apply routeGreenhouse
Salary not disclosed