IT Data Analytics Specialist - Data Engineer
Job Description
Design and deliver governed, reusable data products by building reliable ETL/ELT pipelines on AWS and Snowflake.
Responsibilities
- Design, build, and maintain scalable ETL/ELT pipelines using Matillion on AWS to ingest, transform, and curate data in Snowflake.
- Implement data layer patterns across raw/staging, silver (cleansed and conformed), and gold (business-ready and aggregated) datasets.
- Develop and publish certified Data Products that are governed, documented, reusable, and available via the Data Marketplace for self-service use.
- Build and optimize Snowflake objects, including warehouses, schemas, tables, views, streams, tasks, and stored procedures.
- Improve Snowflake performance using clustering, caching, micro-partitioning, and resource monitoring.
- Support Data Marketplace design by ensuring Data Products are cataloged, discoverable, and aligned to enterprise data governance standards.
- Define and enforce data contracts, SLAs, and quality standards for each published Data Product.
- Define and automate data validation, profiling, and quality checks embedded in ETL pipelines to protect Data Product integrity.
- Collaborate with data governance to establish naming standards, lineage tracking, metadata management, and access policies.
- Ensure end-to-end data lineage from source systems through bronze/silver/gold layers to the final Data Product.
- Partner with data analysts, data scientists, AI/ML teams, product teams, and business stakeholders to translate requirements into Data Products.
- Work with the Data Marketplace product owner on the Data Product lifecycle: creation, certification, publication, and deprecation.
- Lead code reviews, mentor team members, and promote engineering best practices.
- Contribute to architectural decisions and participate in strategic data roadmap planning.
- Build CI/CD pipelines for Matillion job deployments and Snowflake object management.
- Create reusable ETL frameworks, templates, and patterns to speed up new Data Product delivery.
- Implement monitoring, alerting, and operational excellence practices for pipeline health.
- Automate Data Product certification workflows, including quality gates and approval processes.
Requirements
- 5+ years of experience as a Data Engineer or similar role.
- Strong expertise with Snowflake, including SQL, performance tuning, data modeling, warehouse configuration, streams, and tasks.
- Hands-on experience with Matillion on AWS; comparable ETL/ELT tools include Talend, Informatica, dbt, or Azure Data Factory.
- Working understanding of Data Product and Data Marketplace concepts, including data contracts, data-as-a-product, and self-service data consumption.
- Advanced SQL skills for complex query development and optimization.
- Solid data modeling knowledge, including dimensional models, star schemas, and data warehouse/data lake concepts with bronze/silver/gold architecture.
- Cloud experience with AWS.
- Proficiency in Python, Airflow, or other orchestration tools.
- Familiarity with Git, CI/CD, and general DevOps practices.
- Knowledge of data governance principles, metadata management, and data cataloging tools.
Technologies
- Snowflake, SQL, Matillion, AWS, Talend, Informatica, dbt, Azure Data Factory
- Python, Airflow, Git, CI/CD
Preferred Qualifications
- Experience building or contributing to an enterprise Data Marketplace or Data Catalog.
- Familiarity with data mesh or data product thinking methodologies.
- Experience with near real-time data ingestion tools such as OpenFlow, Kafka, or CDC.
- Exposure to AI/ML data preparation and delivering AI-enabled datasets.
- Experience with data lineage and impact analysis tooling.
Location: Westerville, OH (onsite)