This position is no longer accepting applications
Closed on August 4, 2026.
This role is filled — get an email when new SQL roles open on DataJobs.io:
AWS Cloud Data Engineer
Amazon Emr
Amazon Web Services
AWS
Aws Glue
Big Data
Bigdata
Cloud
Cloud Platforms
Data Architecture
Data Engineer
Data Integration
Data Lake
Data Pipeline
Data Pipelines
Data Platform
Data Processing
Data Warehouse
Database
Delta
Dynamodb
EMR
Engineer
ETL
Information Technology (IT)
SQL
View similar jobs
Get alerted when similar jobs are posted — set up a New SQL jobs on DataJobs.io alert.
See other roles at Guidehouse.
Job Description
Benefits
- Medical, Rx, Dental & Vision Insurance
- Personal and Family Sick Time & Company Paid Holidays
- Parental Leave
- 401(k) Retirement Plan
- Group Term Life and Travel Assistance
- Voluntary Life and AD&D Insurance
- Health Savings Account, Health Care & Dependent Care Flexible Spending Accounts
- Transit and Parking Commuter Benefits
- Short-Term & Long-Term Disability
- Tuition Reimbursement, Personal Development, Certifications & Learning Opportunities
- Employee Referral Program
- Corporate Sponsored Events & Community Outreach
- Care.com annual membership
- Employee Assistance Program
- Supplemental Benefits via Corestream (Critical Care, Hospital Indemnity, Accident Insurance, Legal Assistance and ID theft protection, etc.)
- Position may be eligible for a discretionary variable incentive bonus
Overview
Based in San Antonio, this onsite role centers on designing, building, and maintaining data pipelines and processing solutions within the AWS ecosystem. The position involves partnering with database administrators, data scientists, and business stakeholders to deliver scalable data capabilities that drive informed decisions.
Responsibilities
- Design and sustain end-to-end data pipelines and processing solutions in the AWS cloud.
- Collaborate with the AWS Cloud DBA to migrate data from legacy applications to the AWS platform, ensuring performance and reliability.
- Apply hands-on experience with AWS data services and modern data architecture patterns to deliver scalable solutions with cross-functional teams.
- Architect and implement scalable pipelines using AWS services such as AWS Glue, Step Functions, Lambda, and Kinesis.
- Build ETL and ELT processes to ingest, transform, and load data into data warehouses and data lakes from diverse sources.
- Design and maintain data lakes on S3, implement data cataloging with AWS Glue Catalog, and optimize storage formats like Parquet and Delta.
- Develop data warehouse solutions with Redshift and integrate them with existing RDS/Aurora databases managed by the DBA team.
- Develop real-time and batch processing solutions using Kinesis Data Streams, Kinesis Analytics, EMR, and AWS Batch.
- Create and maintain data models, schemas, and documentation.
- Work with the DBA team to optimize data access and query performance across relational and analytical databases.
- Establish automated data quality checks, monitoring, and alerting systems.
- Implement data governance policies and ensure compliance with data retention, privacy, and security requirements.
Requirements
- US Citizenship or Green Card is required.
- Must be able to OBTAIN and MAINTAIN a Federal or DoD Public Trust; candidates must obtain approved adjudication prior to onboarding with Guidehouse. ACTIVE PUBLIC TRUST or SUITABILITY is preferred.
- Bachelor’s degree in computer science, Data Engineering, or a related field; four years of additional experience may substitute for the degree.
- Minimum six years of data engineering experience, with at least three years in AWS cloud environments.
- Strong expertise with AWS data services including S3, Glue, DMS, Athena, Redshift, EMR, Kinesis, and Lambda.
- Proficiency in Python, Scala, or Java for data processing and pipeline development.
- Experience with SQL and working knowledge of relational databases (PostgreSQL, Oracle) and NoSQL systems (DynamoDB, DocumentDB).
- Understanding of data modeling for both transactional and analytical workloads.
- Experience with Infrastructure as Code tools (Terraform, CloudFormation, CDK) and CI/CD pipelines for data engineering workflows.
- Strong analytical and problem-solving abilities with a focus on data quality and system reliability.
- Proven ability to collaborate effectively with DBAs, data scientists, and business stakeholders.
Technologies
- S3, Glue, DMS, Athena, Redshift, EMR, Kinesis, Lambda
- Python, Scala, Java
- SQL, PostgreSQL, Oracle, DynamoDB, DocumentDB
- Terraform, CloudFormation, CDK
- Kinesis Data Streams, Kinesis Analytics, AWS Batch
- Parquet, Delta, AWS Glue Catalog, RDS, Aurora
- Git, Docker, Kubernetes, ECS, EKS
- QuickSight, Tableau, PowerBI
- JSON, Avro, ORC
- Apache Spark, Apache Airflow
Travel
Up to 10 percent travel may be required.
Clearance
Ability to Obtain Public Trust.