xelys jobs xelys jobs

Databricks Data Engineer

Guidehouse

hybridmidpermanentbackenddata United States 57 days ago via LinkedIn

See how well this job matches your profile

Sign up to get an AI match score and generate a tailored application in seconds.

Get your match score

Tags

DatabricksPySparkSQLDelta LakeUnity CatalogETL/ELTBatch and Streaming PipelinesData GovernanceData QualitySpark

About the role

Role Overview

Databricks Data Engineer on Guidehouse’s AI & Data team, supporting client engagements focused on scalable data pipelines, data transformation, platform modernization, and advanced analytics/AI solution delivery.

Responsibilities

  • Design, build, and maintain scalable data pipelines using Databricks, PySpark, SQL, and Delta Lake
  • Develop and support batch and streaming workflows (ingestion, transformation, validation, publishing of curated data products)
  • Write and optimize Python/PySpark/SQL for data processing, orchestration, and performance tuning
  • Work with large-scale datasets using Databricks/Spark/Delta Lake, Unity Catalog, and cloud storage
  • Troubleshoot pipeline failures, performance issues, data quality problems, and bottlenecks
  • Translate business/technical requirements into data engineering designs and orchestration workflows
  • Perform data validation, quality checks, code reviews, and issue resolution to ensure reliable data products
  • Collaborate with solution architects, data scientists, analysts, DevOps engineers, and client stakeholders
  • Communicate delivery impacts and engineering considerations to technical and non-technical audiences
  • Document pipelines, datasets, models, workflows, and architectural decisions
  • Follow data governance, security, lineage, and compliance standards within the Databricks ecosystem

Requirements

  • Bachelor’s degree in computer science, engineering, mathematics, statistics, or a related field
  • 3–8 years of experience in data engineering, data architecture, or cloud data platform implementation
  • Strong experience with Python, PySpark, and SQL for transformation and pipeline development
  • Experience with Databricks, Spark, Delta Lake, Unity Catalog (or similar cloud-native platforms)
  • Experience building batch/streaming ETL/ELT pipelines and reusable data assets
  • Knowledge of data modeling, data warehousing, data quality validation, and large-scale processing
  • Ability to troubleshoot issues, communicate recommendations clearly, and deliver in team-based environments
  • Experience applying governance, access control, lineage, and security in cloud data platforms

Nice to Have

  • 2+ years hands-on experience with the Databricks platform
  • Additional Unity Catalog experience for governance, schema design, modeling standards, access controls, lineage, and secure enterprise data management

About Guidehouse

Guidehouse is a consulting firm that delivers data, AI, and technology solutions for clients. The role sits within their AI & Data team, supporting client projects that include large-scale data pipelines, analytics, and platform modernization.

Scraped 7/29/2026