Associate Data Engineer
Big Data
Bigdata
Cloud Operations
Cloud Platform
Cloud Platforms
Data
Data Analysis
Data Engineer
Data Engineering
Data Integration
Data Operations
Data Pipeline
Data Pipelines
Data Platform
Data Processing
Data Warehouse
Data Warehousing
Database
Databases
ETL
Informatica
Information Technology (IT)
Integration
Spark
SQL
Job Description
Join the S&S Data Science team’s External Data function by building and operating pipelines that bring third-party information into underwriting, claims, and business intelligence workflows.
Responsibilities
- Build and maintain batch and streaming ETL/ELT pipelines that load into a cloud data warehouse
- Develop, schedule, and monitor cloud data jobs, including failure diagnosis and cost and performance tuning
- Create idempotent load logic to support reprocessing and backfills without duplicate records
- Integrate external data providers and REST APIs with authentication, error handling, and retries
- Implement data quality checks, validation, and alerting to identify pipeline issues early
- Manage and deploy pipelines using infrastructure-as-code and CI/CD
- Support data migrations and service updates with minimal downtime
- Deliver clean, analytics-ready, and model-ready datasets for data scientists and analysts
Requirements
- 1-5+ years of professional experience in data engineering or a related software/data role
- Bachelor’s degree in Computer Science, Data Engineering, or a related technical field (or equivalent experience)
- Strong programming skills in Python and SQL
- Working knowledge of ETL/ELT and building data pipelines
- Experience with at least one major cloud platform, with AWS preferred
- Familiarity with REST APIs and integrating external data sources
- High attention to data quality and reliability
- Ability to communicate clearly with both technical and business audiences
Technologies
- Python, SQL
- AWS, ETL, ELT
- REST APIs
- AWS Glue, AWS Lambda, S3
- Snowflake
- Apache Spark, Kafka, Kinesis
- Infrastructure-as-code, CDK, Terraform, CloudFormation
- CloudWatch, CI/CD (listed as CI CD)
Benefits
- Competitive compensation package (USD 71,600 - 105,000 per year)
- Generous 401K employer match
- Employee Stock Purchase plan with employer matching
- Generous Paid Time Off
- Excellent benefits that go beyond health, dental & vision
- Tuition reimbursement, industry-related certifications, and professional training
- Dynamic, ambitious, fun, and exciting work environment
- Matching donation program, volunteer opportunities, and an employee-driven corporate giving program
Desired Skills (Nice to Have)
- Hands-on experience with AWS data services such as Glue, Lambda, and S3
- Experience loading and modeling data in a cloud data warehouse (e.g. Snowflake)
- Experience with distributed processing frameworks such as Apache Spark
- Experience with streaming technologies such as Kafka or Kinesis
- Experience with infrastructure-as-code (CDK, Terraform, or CloudFormation) and monitoring with CloudWatch
- Knowledge of geospatial analysis
- Familiarity with zero-downtime deployment patterns such as blue-green deployments
- Interest in insurance, underwriting, or claims
Location: Morristown, NJ (remote)