DataJobs.io
← Back to all jobs

Job Description

Amazon's PXT Central Science team is seeking a Data Engineer to design, build, and maintain data pipelines and analytics that reveal the drivers behind shifts in employee sentiment, actions, and business outcomes. This on-site role in Arlington, VA collaborates with economists, data scientists, and software engineers to productionize science models and deliver scalable data solutions that support decision-making across the organization.

Responsibilities

  • Data pipeline development: Design and maintain scalable data pipelines using native AWS services (Glue, EMR, Lambda); implement monitoring and error handling for data workflows; optimize performance, reliability, and cost efficiency.
  • Model productionization and API development: Develop and maintain APIs and data serving layers that productionize science models for downstream consumption; build batch and real-time inference pipelines.
  • Data integration and quality: Build scalable feature extraction and processing frameworks for diverse data types; develop robust data quality and validation checks; create flexible schemas supporting evolving requirements.
  • Cross-team collaboration: Partner with economics, data science, and software engineering teams to translate analytical requirements into production-ready solutions; participate in technical design reviews and architecture discussions.
  • Analytics and infrastructure: Maintain layered data systems used by economists and scientists; build automated reporting solutions; operate across multiple interconnected AWS accounts with security best practices.

Requirements

  • Knowledge of professional software engineering and best practices for the full software development lifecycle, including coding standards, software architectures, code reviews, source control, continuous deployments, testing, and operational excellence.
  • 3+ years of data engineering experience.
  • Experience in at least one modern scripting or programming language, such as Python, Java, Scala, or NodeJS.
  • Experience with data modeling, warehousing, and building ETL pipelines.
  • Experience with AWS technologies like Redshift, S3, AWS Glue, EMR, Kinesis, FireHose, Lambda, and IAM roles and permissions.
  • Experience with non-relational databases / data stores (object storage, document or key-value stores, graph databases, column-family databases).
  • Bachelor's degree or foreign equivalent in computer science, engineering, mathematics or equivalent.

Technologies

  • Python
  • Java
  • Scala
  • NodeJS
  • AWS Glue
  • EMR
  • Lambda
  • Redshift
  • S3
  • Kinesis
  • FireHose
  • IAM
  • Hadoop
  • Hive
  • Spark
  • Generative AI

Benefits

  • Health insurance (medical, dental, vision, prescription)
  • 401(k) matching
  • Paid time off
  • Parental leave
  • Sign-on payments
  • Restricted stock units (RSUs)
  • Adoption and Surrogacy Reimbursement coverage
  • Employee Assistance Program (EAP) and Mental Health Support
  • Medical Advice Line
  • Flexible Spending Accounts (FSA)
  • Basic Life & AD&D insurance
  • Optional supplemental life plans

About the Team

The Central Science Team within Amazon’s People Experience and Technology organization (PXTCS) brings together economics, behavioral science, statistics, machine learning, and Generative AI to proactively identify mechanisms and process improvements that enhance Amazon and the lives, well-being, and value of work for Amazonians. This interdisciplinary group combines the talents of science, engineering, and user experience to develop solutions that measurably advance this goal.

Similar Jobs