DataJobs.io
← Back to all jobs

Job Description

Design, build, and optimize scalable data pipelines and architectures for cloud-based analytics and processing.

Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines using AWS services
  • Develop and optimize batch and real-time data processing systems
  • Ensure data quality, integrity, and governance across platforms and workflows
  • Partner with stakeholders to translate business requirements into technical solutions
  • Implement data integration solutions across multiple sources and formats
  • Monitor and troubleshoot data workflows to support performance and reliability
  • Optimize costs and performance for AWS and GCP data infrastructure
  • Collaborate with data scientists, analysts, and DevOps teams

Requirements

  • Bachelor’s degree in Computer Science, Engineering, or a related field
  • 3-8+ years of experience in data engineering or related roles
  • Proficiency in SQL and at least one programming language: Python, Scala, or Java
  • Experience with ETL tools and frameworks
  • Solid understanding of data modeling, warehousing, and big data concepts
  • Familiarity with distributed processing frameworks: Spark, Hadoop
  • Experience with CI/CD pipelines and version control (Git)

Technologies

  • AWS, GCP
  • Databricks, Snowflake
  • dbt, SQL
  • Python, Scala, Java
  • ETL tools
  • Spark, Hadoop
  • Git, CI/CD pipelines
  • Kafka
  • Terraform, CloudFormation
  • Power BI, Tableau

Preferred qualifications

  • Experience with Apache Spark, Kafka, or Databricks
  • Knowledge of data governance, security, and compliance
  • Familiarity with infrastructure as code (Terraform, CloudFormation)
  • AWS certifications (e.g., AWS Certified Data Engineer / Solutions Architect)
  • Experience with BI tools (Power BI, Tableau)

Key competencies

  • Strong problem-solving and analytical skills
  • Excellent communication and collaboration abilities
  • Attention to detail and commitment to data quality
  • Ability to work in a fast-paced, agile environment

Nice to have

  • Experience with machine learning data pipelines
  • Knowledge of real-time analytics and streaming architectures
  • Exposure to multi-cloud or hybrid environments

Benefits

  • Paid time off based on employee grade (A-F), defined by policy: Vacation 12-25 days depending on grade, Company paid holidays, Personal Days, Sick Leave
  • Medical, dental, and vision coverage (or provincial healthcare coordination in Canada)
  • Retirement savings plans (e.g., 401(k) in the U.S., RRSP in Canada)
  • Life and disability insurance
  • Employee assistance programs
  • Other benefits as provided by local policy and eligibility

Location: New York, NY (onsite) | Salary: USD 70,000 - 95,000 per yearly | Minimum experience: 3 years

Similar Jobs