Databricks Data Engineer
Guidehouse
hybridmidpermanentbackenddata United States Yesterday via LinkedIn
See how well this job matches your profile
Sign up to get an AI match score and generate a tailored application in seconds.
Get your match scoreTags
DatabricksPySparkSQLDelta LakeUnity CatalogSparkETL/ELTData ModelingData GovernanceStreaming Data Pipelines
About the role
Role Overview
Databricks Data Engineer to join Guidehouse’s AI & Data team. You will design, build, and maintain scalable, secure data pipelines and curated data products for client projects, using Databricks and related cloud-native data technologies.
What You Will Do
- Design, build, and maintain scalable data pipelines using Databricks, PySpark, SQL, Delta Lake, and cloud-native tooling.
- Develop and support batch and streaming workflows (ingestion → transformation → validation → publishing of data products).
- Write, optimize, and maintain Python/PySpark/SQL for data processing, orchestration, and performance tuning.
- Work with large-scale datasets using Databricks, Spark, Delta Lake, Unity Catalog, and cloud storage services.
- Troubleshoot pipeline failures, performance issues, data quality problems, and workflow bottlenecks.
- Translate business/technical requirements into data engineering designs, processing scripts, and orchestration workflows.
- Perform data validation, quality checks, code reviews, and resolve issues to ensure reliable data products.
- Collaborate with solution architects, data scientists/analysts, DevOps engineers, and client stakeholders.
- Communicate technical concepts and delivery impacts to both technical and non-technical audiences.
- Document pipelines, datasets, data models, workflows, and engineering decisions.
- Follow Databricks governance, security, lineage, and compliance standards.
What You Will Need
- Bachelor’s degree in CS/engineering/math/stats or related field.
- 3–8 years of relevant experience in data engineering, data architecture, or cloud data platform implementation.
- Strong experience with Python, PySpark, and SQL for transformations and pipeline development.
- Experience with Databricks, Spark, Delta Lake, Unity Catalog, or similar cloud-native platforms.
- Experience building batch/streaming ETL/ELT and reusable data assets.
- Knowledge of data modeling, warehousing, data quality validation, and large-scale processing concepts.
- Ability to troubleshoot, clearly communicate recommendations, and work in team-based delivery.
- Experience with data governance, access control, lineage, and security in cloud data platforms.
Nice to Have
- 2+ years hands-on Databricks experience.
- Experience with Unity Catalog for governance, schema design, data modeling standards, access controls, lineage, and secure enterprise data management.
About Guidehouse
Guidehouse is a consulting firm that delivers AI and data solutions and supports clients with large-scale technology and analytics projects. The role sits within the company’s AI & Data team, focusing on building secure, reliable data engineering platforms and advanced analytics/AI solution delivery.
Scraped 8/4/2026