Skip to main content
Muttdata

Data Engineer - Microsoft Fabric

RemoteArgentina onlyArchived
Published
Role
Data Engineering
Experience
Senior
Employment
Full-time
Company size
Startup
Salary not disclosed
Check eligibility

Open to AR only. Set where you work from to check your eligibility.

No BS summary

Senior Data Engineer with 4+ years of production experience in data pipelines. Must have deep PySpark and Delta Lake experience, hands-on Microsoft Fabric (lakehouses, notebooks, Data Pipelines, OneLake shortcuts), and strong SQL. Needs familiarity with Azure ecosystem (Blob Storage, Entra ID, Key Vault, Static Web Apps). Must be comfortable owning end-to-end pipeline reliability and working effectively in a distributed team with overlap to North American business hours. Strong written and spoken English required.

Core skills

PySparkDelta LakeMicrosoft Fabric

Required skills

lakehousesnotebooksData PipelinesOneLake shortcutsSQLwindow functionsCTEsanalytical patternspartition pruningpredicate pushdownBlob StorageEntra IDKey VaultStatic Web Apps

Required languages

English

What you'll do

  • Own end-to-end pipeline reliability across Bronze, Silver, and Gold layers (PySpark + Delta Lake on Microsoft Fabric).
  • Build and maintain ingestion notebooks for new retailers and syndicated data partners as we scale.
  • Harden existing pipelines - idempotent replaceWhere patterns, partition strategies, ZORDER optimization, schema enforcement, and dedup logic.
  • Run and improve our daily and weekly orchestration through Fabric Data Pipelines and scheduled notebook runs.
  • Diagnose and resolve runtime issues across the medallion stack - including the parts of Fabric that don’t behave the way the docs say they do.
  • Manage lakehouse shortcuts, Azure Blob Storage accounts (for non-HNS sources), and data source authentication via Entra / Key Vault.
  • Partner with the AI engineering team to keep the Gold layer clean and queryable for our RAG, reporting, and attribution use cases.
  • Contribute to data modeling decisions across our unified retailer schemas, keeping naming conventions consistent across the platform.
  • Document patterns so the next engineer can move at our pace.

What they require

  • 4+ years building data pipelines in production.
  • Deep PySpark and Delta Lake experience - you’ve shipped real medallion architectures, not just read about them.
  • Hands-on Microsoft Fabric experience: lakehouses, notebooks, Data Pipelines, OneLake shortcuts. If you’ve worked through Fabric’s quirks (cells not reliably sharing Python variables, shortcut type limitations on non-HNS storage, etc.), that’s exactly the experience we want.
  • Strong SQL, including window functions, CTEs, and analytical patterns. Comfortable reasoning about partition pruning and predicate pushdown.
  • Comfortable owning end-to-end pipeline reliability - not just writing the happy path. You think about reruns, backfills, late-arriving data, and what happens at 4am when something breaks.
  • Azure ecosystem familiarity: Blob Storage, Entra ID, Key Vault, Static Web Apps.
  • Declarative, sparse code style. You prefer fixing the schema over patching the symptom.
  • Strong written and spoken English; able to work effectively in a distributed team with overlap to North American business hours.

Benefits

  • Remote-first culture – work from anywhere! 🌍
  • AWS, DBT, Google Cloud, Azure & Databricks certifications fully covered
  • In-Company English Lessons.
  • Birthday off + an extra vacation week (Mutt Week! 🏖️)
  • Referral bonuses – help us grow the team & get rewarded!
  • Maslow: Monthly credits to spend in our benefits marketplace.
  • ✈️🏝️ Annual Mutters' Trip – an unforgettable getaway with the team!

At Muttdata, we build innovative Data Products and Machine Learning solutions that help companies solve complex business challenges. As a fast-growing, remote-first startup, we're passionate about technology, collaboration, and continuous learning.

🇦🇷 ArgentinaAIStartup
Salary not disclosed