xelys jobs xelys jobs

Data Platform Engineer

Bright Vision Technologies

full-remoteseniorpermanentbackenddata Cary, NC 64 days ago via LinkedIn
100,000 - 150,000 USD/annual

See how well this job matches your profile

Sign up to get an AI match score and generate a tailored application in seconds.

Get your match score

Tags

Apache SparkHadoopKafkaHiveHDFSSqoopHBaseFlinkAirflowETL/ELT

About the role

Role Overview

Bright Vision Technologies is seeking a Data Platform Engineer to design, build, and operate large-scale data processing pipelines and analytics platforms using Hadoop and related big-data technologies. This is a 100% remote (U.S.), full-time role.

Responsibilities

  • Design, develop, and operate end-to-end big-data pipelines across relational, file-based, streaming, and API-driven sources.
  • Build robust ETL/ELT workflows with Apache Spark, Hive, Pig, and Sqoop, with focus on data quality, idempotency, error handling, and recoverability.
  • Create high-throughput streaming pipelines using Kafka, Spark Streaming, or Flink; integrate with downstream systems.
  • Optimize Spark and MapReduce jobs (partitioning, memory, serialization, skew) to meet SLAs at minimal cost.
  • Design and maintain data models and storage on HDFS, Hive, HBase, and lakehouse formats such as Parquet/ORC and Delta, Iceberg, Hudi.
  • Implement data governance, lineage, and quality controls in collaboration with governance/security teams.
  • Set up monitoring, alerting, and logging for pipeline reliability and proactive failure detection.
  • Partner with data scientists and analysts to deliver curated, reliable, well-documented datasets.
  • Automate pipeline orchestration using Airflow, Oozie, or similar workflow engines.
  • Continuously evaluate and adopt new big-data/cloud technologies (e.g., EMR, Databricks, Snowflake, BigQuery) when they provide meaningful improvements.
  • Lead performance reviews and architecture audits, proposing refactoring and optimization.
  • Document architectures, schemas, pipeline behavior, and runbooks; support scalability.
  • Mentor junior engineers and contribute to engineering standards and best practices.

Required Qualifications

  • Bachelor’s degree in Computer Science, Engineering, or a related technical field.
  • 5+ years experience designing and operating big-data pipelines on Hadoop.
  • Strong hands-on expertise with Apache Spark (Scala, Python, or Java) in production.
  • Solid experience with Hive, HDFS, Sqoop, HBase, and the broader Hadoop ecosystem.
  • Hands-on streaming experience with Kafka, Spark Streaming, or Flink.
  • Strong SQL skills; experience with both relational and NoSQL data stores.
  • Experience with Airflow or Oozie for orchestration.
  • Strong distributed systems understanding (partitioning, replication, fault tolerance).

Nice-to-Haves / Additional Signals

  • Experience integrating modern lakehouse platforms and formats (e.g., Delta/Iceberg/Hudi).
  • Experience with cloud big-data ecosystems such as EMR or Databricks.
  • Leadership/mentoring experience for junior engineers.

About Bright Vision Technologies

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions. It supports clients across the United States with engineering services and delivery of data-driven platforms.

Scraped 7/23/2026