DataJobs.io
← Back to all jobs

Job Description

Vytwo is seeking a Senior Data Engineer to design and maintain scalable data pipelines and improve Databricks and Spark workloads so downstream analytics teams receive reliable, analytics-ready datasets. This role is based in Minneapolis, MN with onsite work.

Key Responsibilities

  • Design, develop, and maintain scalable data pipelines using Apache Spark with PySpark and/or Scala.
  • Build and optimize data workflows on Databricks, including Delta Lake, notebooks, and scheduled jobs.
  • Ingest, transform, and curate large-scale structured and semi-structured datasets.
  • Conduct performance tuning and cost optimization for Spark workloads and Databricks clusters.
  • Implement data quality checks, monitoring, and error handling practices.
  • Collaborate with analytics and business stakeholders to deliver well-modeled, analytics-ready data.
  • Support batch processing and, where applicable, streaming data pipelines.
  • Apply best practices for testing, documentation, security, and version control.

Required Qualifications

  • 5+ years of experience in Data Engineering or a related role.
  • Strong hands-on experience with Apache Spark (PySpark or Scala).
  • Proven experience working in Databricks environments.
  • Strong SQL skills, including experience with relational and analytical databases.
  • Experience building and maintaining ETL/ELT pipelines at scale.
  • Familiarity with modern data lake architectures; Delta Lake preferred.
  • Experience using Git-based version control.
  • Must be a US Citizen.

Technology Stack

  • Apache Spark, PySpark, Scala
  • Databricks, Delta Lake
  • Git, SQL

Location and Eligibility

Minneapolis, MN (onsite). Must be currently located in Minnesota (MN), or willing to relocate. Must be a U.S. Citizen.

Benefits

  • Flexible work from home options available.

Similar Jobs