Skip to main content

Data Platform Engineer

RemoteLATAM
Published
Role
Data Engineering
Employment
Contract
Salary not disclosed
Check eligibility

Open to Anywhere in LATAM. Set where you work from to check your eligibility.

No BS summary

Senior data platform engineer to migrate a large Azure Data Factory estate (2,094 pipelines, 471 Spark dataflows) to Databricks on AWS. Must have production Databricks, Spark/PySpark, Python, ADF and AWS data services. Full-time remote, open to candidates across LATAM.

Core skills

DatabricksPySpark

Required skills

SparkPythonAzure Data FactoryAWS S3AWS GlueAWS AthenaAWS LambdaAWS Step FunctionsSQL

Optional skills

Delta LakeUnity CatalogAzure SynapseTerraformAirflowdbtKafkadata modeling

Required languages

English professional

What you'll do

  • Convert Azure Data Factory pipelines into Databricks workflows on AWS, building reusable templates rather than migrating one at a time
  • Rehost Databricks workspaces onto AWS and migrate ADLS Gen2 storage to S3
  • Rewrite ADF Web Activities as Lambda functions or Step Functions tasks, and replace ADF-specific scaling with native Databricks mechanisms
  • Build and tune PySpark transformations for production data volumes
  • Replace Azure Synapse Serverless with Databricks SQL Warehouse
  • Reconcile migrated data against source systems as part of the definition of done

What they require

  • Production experience with Databricks: workspaces, jobs and workflows
  • Strong Spark and PySpark experience for real data volumes, including tuning
  • Production-grade Python
  • Experience building or migrating Azure Data Factory pipelines, with a solid understanding of the ADF activity model
  • AWS data services: S3, Glue, Athena, Lambda and Step Functions
  • Advanced SQL, including reading and reasoning about stored procedures
  • Professional written and spoken English
Salary not disclosed