Project - Data Engineer
Job Description
Deloitte is seeking a Project Delivery Data Engineer in Stamford, CT on site to build data pipelines and analytics-ready datasets in AWS and Snowflake for a large onshore/offshore program.
Responsibilities
- Design and optimize AWS data pipelines with Python to ingest, transform, and deliver data to Snowflake and downstream consumers.
- Create and maintain Snowflake objects (schemas, tables, views) and implement performant SQL transformations to produce curated, analytics-ready datasets.
- Implement workflow automation and scheduling using Airflow/MWAA, Step Functions, or Glue, including dependencies, retries, and logging.
- Embed data quality checks and basic observability with validation rules, reconciliations, and alerts; support incident triage and remediation.
- Improve pipeline and query performance through guidance on efficient Python practices, S3 partitioning/file formats (Parquet/Delta), and Snowflake warehouse usage and tuning.
- Follow CI/CD and IaC standards with Git workflows and Terraform/CloudFormation changes to promote code across environments.
- Collaborate with analysts, product owners, and source-system teams to clarify requirements and validate outputs; participate in sprint ceremonies and estimations.
- Contribute to code reviews, unit tests, and peer debugging; adhere to team engineering standards.
- Maintain regular communication with Engagement Managers, project team members, and cross-functional stakeholders; escalate matters requiring engagement management input.
- Lead client engagement workstreams focused on improvement, optimization, and transformation of processes, implementing best-practice workflows and driving operational outcomes.
Requirements
- Minimum 1 year of experience building or enhancing data pipelines and curated datasets for analytics and downstream consumers.
- At least 1 year of hands-on SQL and Python experience, including Snowflake and/or PySpark for transformations and scalable processing.
- 1+ year of cloud data engineering experience on AWS (preferred) or Azure/GCP, including orchestration or scheduling (Airflow/MWAA, Step Functions, Glue, ADF/Fabric Data Factory).
- Understanding of ELT patterns and lakehouse/warehouse concepts; familiarity with S3 file formats and partitioning (Parquet, Delta).
- Working knowledge of DevOps practices (Git-based workflows, CI/CD) and exposure to Infrastructure-as-Code (Terraform or CloudFormation).
- Understanding of data quality, basic observability, and metadata/governance fundamentals.
- Bachelor's degree in Computer Science, Information Technology, Computer Engineering, or related IT discipline, or equivalent experience.
- Limited immigration sponsorship may be available.
- Ability to travel approximately 10 percent, depending on assignments and client needs.
- Experience delivering in Agile environments.
- Analytical ability to manage multiple projects and prioritize tasks into manageable work products.
- Ability to work independently or with minimal supervision.
- Excellent written and verbal communication skills.
- Ability to deliver technical demonstrations.
Technologies
- AWS, Python, Snowflake, SQL, PySpark
- Airflow (MWAA), Step Functions, Glue
- Terraform, CloudFormation, Git
- Parquet, Delta, S3
- Azure Data Factory, Fabric Data Factory
Benefits
- Discretionary annual incentive program