DataJobs.io
← Back to all jobs

Job Description

Design and deliver governed, reusable data products by building reliable ETL/ELT pipelines on AWS and Snowflake.

Responsibilities

  • Design, build, and maintain scalable ETL/ELT pipelines using Matillion on AWS to ingest, transform, and curate data in Snowflake.
  • Implement data layer patterns across raw/staging, silver (cleansed and conformed), and gold (business-ready and aggregated) datasets.
  • Develop and publish certified Data Products that are governed, documented, reusable, and available via the Data Marketplace for self-service use.
  • Build and optimize Snowflake objects, including warehouses, schemas, tables, views, streams, tasks, and stored procedures.
  • Improve Snowflake performance using clustering, caching, micro-partitioning, and resource monitoring.
  • Support Data Marketplace design by ensuring Data Products are cataloged, discoverable, and aligned to enterprise data governance standards.
  • Define and enforce data contracts, SLAs, and quality standards for each published Data Product.
  • Define and automate data validation, profiling, and quality checks embedded in ETL pipelines to protect Data Product integrity.
  • Collaborate with data governance to establish naming standards, lineage tracking, metadata management, and access policies.
  • Ensure end-to-end data lineage from source systems through bronze/silver/gold layers to the final Data Product.
  • Partner with data analysts, data scientists, AI/ML teams, product teams, and business stakeholders to translate requirements into Data Products.
  • Work with the Data Marketplace product owner on the Data Product lifecycle: creation, certification, publication, and deprecation.
  • Lead code reviews, mentor team members, and promote engineering best practices.
  • Contribute to architectural decisions and participate in strategic data roadmap planning.
  • Build CI/CD pipelines for Matillion job deployments and Snowflake object management.
  • Create reusable ETL frameworks, templates, and patterns to speed up new Data Product delivery.
  • Implement monitoring, alerting, and operational excellence practices for pipeline health.
  • Automate Data Product certification workflows, including quality gates and approval processes.

Requirements

  • 5+ years of experience as a Data Engineer or similar role.
  • Strong expertise with Snowflake, including SQL, performance tuning, data modeling, warehouse configuration, streams, and tasks.
  • Hands-on experience with Matillion on AWS; comparable ETL/ELT tools include Talend, Informatica, dbt, or Azure Data Factory.
  • Working understanding of Data Product and Data Marketplace concepts, including data contracts, data-as-a-product, and self-service data consumption.
  • Advanced SQL skills for complex query development and optimization.
  • Solid data modeling knowledge, including dimensional models, star schemas, and data warehouse/data lake concepts with bronze/silver/gold architecture.
  • Cloud experience with AWS.
  • Proficiency in Python, Airflow, or other orchestration tools.
  • Familiarity with Git, CI/CD, and general DevOps practices.
  • Knowledge of data governance principles, metadata management, and data cataloging tools.

Technologies

  • Snowflake, SQL, Matillion, AWS, Talend, Informatica, dbt, Azure Data Factory
  • Python, Airflow, Git, CI/CD

Preferred Qualifications

  • Experience building or contributing to an enterprise Data Marketplace or Data Catalog.
  • Familiarity with data mesh or data product thinking methodologies.
  • Experience with near real-time data ingestion tools such as OpenFlow, Kafka, or CDC.
  • Exposure to AI/ML data preparation and delivering AI-enabled datasets.
  • Experience with data lineage and impact analysis tooling.

Location: Westerville, OH (onsite)

Similar Jobs