Skip to main content
Intetics Inc.

Senior Data Engineer - Databricks

RemotePoland, United Kingdom, Germany +2 more only
Published
Role
Data Engineering
Experience
Senior
Employment
Full-time
Salary not disclosed
Check eligibility

Open to PL, GB, DE, ES, RO only. Set where you work from to check your eligibility.

No BS summary

Senior Data Engineer with 4+ years of experience, including at least 2 years with Databricks/Spark on Azure/AWS. Must be proficient in PySpark, SQL, and Python, with hands-on experience in Delta Lake, pipeline performance tuning, and PostgreSQL. Experience with large-scale multi-tenant architectures and data governance is required.

Core skills

DatabricksPySparkDelta Lake

Required skills

Apache SparkSQLPythonPostgreSQLETLELTCI/CDversion controlautomated testingdeployment tooling

Optional skills

Microsoft SQL ServerML feature engineeringfeature storesDatabricks Feature StoreFeastcustomer onboarding automationInfrastructure as Code (IaC)Databricks Certified Data Engineer Associate

Required languages

English

What you'll do

  • Own Databricks production support for the company's data platform, including monitoring, alerting, and incident response across all production data flows.
  • Maintain and report on SLA performance metrics for data pipeline delivery, ensuring visibility into platform health and accountability across internal and external stakeholders.
  • Identify and implement pipeline optimizations that reduce Databricks compute costs, improve throughput, and reduce processing windows while tracking impacts through measurable KPIs.
  • Migrate legacy ETL/ELT pipelines to Databricks, building automation tooling to reduce manual intervention and ensure uninterrupted data delivery during transitions.
  • Support new customer onboarding by provisioning, validating, and hardening tenant data pipelines that deliver reliable, isolated data from day one.
  • Design and build high-performance Databricks pipelines that ingest, transform, and serve ERP and CRM data at scale across both Azure and AWS environments.
  • Own the Delta Lake architecture, including schema design, partitioning strategies, data quality enforcement, and incremental processing patterns.
  • Enforce data security best practices across Databricks environments, including role-based access control, secrets management, and compliance requirements for enterprise business data.
  • Implement data quality monitoring and observability across pipeline health and ML model inputs, ensuring data integrity that directly supports predictive analytics.
  • Apply and enforce multi-tenant data isolation patterns, ensuring reliable and secure data delivery across enterprise customers.
  • Partner with the Enterprise Architecture team to ensure data pipelines integrate seamlessly with the broader AI and analytics ecosystem.
  • Support a globally distributed operation through on-call rotation and after-hours incident response, meeting SLAs across multiple time zones.
  • Maintain technical documentation, runbooks, and architectural decision records, contributing to team knowledge sharing and operational readiness across on-call and incident response scenarios.
  • Apply CI/CD best practices to data pipeline development, including version control, automated testing, and deployment tooling to ensure reliable and repeatable pipeline delivery.

What they require

  • 4+ years of data engineering experience.
  • At least 2 years of experience with Databricks or the Apache Spark ecosystem across Azure and/or AWS.
  • Proficiency in PySpark, SQL, and Python with a strong track record of building and operating production-grade pipelines under SLA constraints.
  • Hands-on experience with Delta Lake, including schema evolution, ACID transactions, optimize/vacuum lifecycle, and both incremental and streaming processing patterns.
  • Hands-on experience with pipeline performance tuning and compute optimization in production Databricks environments.
  • Solid working knowledge of PostgreSQL, including query optimization, schema design, and use as a source or sink in production data pipelines.
  • Experience supporting and maintaining legacy ETL tooling (SSIS, Informatica, custom Python/SQL pipelines, or similar) in production.
  • Experience supporting large-scale multi-tenant architectures with a focus on tenant isolation, per-tenant performance, and data privacy, including navigating tools and platforms that default to single-tenant assumptions.
  • Proven ability to work collaboratively across data science, product, and infrastructure teams, owning end-to-end delivery in a cross-functional environment.
  • Strong understanding of data governance, security, and compliance principles, including access control, data privacy, and protection of sensitive enterprise data across multi-tenant environments.
  • Experience operating Databricks workspaces across both Azure and AWS, including cost governance, cluster management, and cross-cloud data access.
  • Experience optimizing Databricks workloads in a Serverless environment, including compute cost governance and performance tuning for serverless compute.
  • Experience with Microsoft SQL Server in a data engineering or ETL context.
  • Exposure to ML feature engineering or feature stores (Databricks Feature Store, Feast, or similar) supporting predictive analytics.
  • Experience with customer onboarding automation or Infrastructure as Code (IaC) patterns for provisioning tenant data pipelines at scale.

Intetics Inc., a global technology company providing custom software application development, distributed professional teams, software product quality assessment, and “all-things-digital” solutions.

🇺🇸 United StatesCybersecurity

What people say about this company

4.5/ 5

Salary not disclosed