xelys jobs xelys jobs

AI Data Scientist

DeepRunner AI

hybridmidpermanentdata Silicon Valley, CA 27 days ago via LinkedIn

See how well this job matches your profile

Sign up to get an AI match score and generate a tailored application in seconds.

Get your match score

Tags

AI Data ScienceSynthetic DataGANsVAEsPythonPandasNumPyScikit-learnData PipelinesApache Airflow

About the role

Role overview

DeepRunner AI is seeking an AI Data Scientist to own the full data lifecycle—from acquisition and cleaning to preparation for training cutting-edge AI models. A key part of the role is creating synthetic data generation techniques to improve model performance and augment existing datasets.

Responsibilities

  • Manage the entire data lifecycle: acquisition, cleaning, processing, and preparation for training
  • Develop synthetic data generation methods using generative models (GANs, VAEs)
  • Design and implement robust data pipelines using orchestration tools
  • Perform data acquisition via web scraping, API integration, and database querying
  • Execute data cleaning, normalization, and preprocessing for structured, unstructured, and time-series data
  • Implement data quality monitoring, validation, and governance
  • Create data visualizations and support feature engineering

Requirements

  • 3+ years of data science experience, focused on data preparation and feature engineering
  • Strong data acquisition skills: web scraping, APIs, SQL/NoSQL querying
  • Solid hands-on work with data cleaning/normalization/preprocessing across multiple data types
  • Proven synthetic data generation experience using GANs and VAEs
  • Understanding of data pipeline architectures and orchestration tools: Apache Airflow, Luigi, or Prefect
  • Proficiency in Python and libraries such as Pandas, NumPy, and scikit-learn
  • Experience with cloud storage/processing platforms such as AWS S3, Google Cloud Storage, or Azure Blob Storage

Nice to have

  • Bachelor’s degree in CS/Statistics/Math or related
  • Experience with big data (Spark, Hadoop)
  • Data quality monitoring/validation and data governance/privacy experience
  • Familiarity with feature stores and feature engineering platforms
  • Experience with Tableau or Power BI

Benefits / what we offer

  • Competitive salary with equity participation
  • Comprehensive health coverage (medical/dental/vision)
  • Flexible work arrangements and remote options
  • Professional development and conference attendance
  • Work with cutting-edge data and AI technology

About DeepRunner AI

DeepRunner AI is an AI-focused company building AI-driven automation solutions. The role emphasizes end-to-end data engineering and AI data science to support training and deployment of advanced models.

Scraped 7/3/2026