Senior Data Engineer
Senior
3d Design Tools
Amazon Rds Aurora
Amazon Web Services
AWS
Aws Athena
Aws Data Warehouse
Aws Glue
Big Data
Cloud
Cloud Computing
Cloud Data Engineering
Cloud Data Warehouse
Cloud Operations
Cloud Platforms
Data
Data Analysis
Data Analytics
Data Architecture
Data Engineer
Data Engineering
Data Integration
Data Lake
Data Pipeline
Data Platform
Data Processing
Data Warehouse
Database
Databases
ETL
Informatica
IT Services
Programming Language
Programming Languages
Rendering Engines
Reporting and Analytics
SQL
Job Description
Vertage is building NextBrain, a cloud data foundation meant to securely and reliably power access to operational, engineering, analytical, and contextual information. As a Senior Data Engineer in the United States (onsite), you will design and deliver the core pipelines, storage architecture, data models, APIs, and automated data-quality controls on AWS.
This role focuses on making NextBrain data dependable for both applications and AI capabilities, with careful attention to governance, lineage, access control, and the performance and cost of production workloads.
What you will do
- Design and implement the NextBrain cloud data architecture.
- Build scalable ingestion, transformation, storage, and serving pipelines on AWS.
- Create pipelines for operational, historical, engineering, application, and contextual data.
- Design data models optimized for analytical applications and AI consumption.
- Define storage patterns across relational, object, time-series, and analytical data stores.
- Develop APIs and services that let NextBrain tools and agents retrieve data consistently.
- Implement automated data-quality validation, reconciliation, and monitoring.
- Establish metadata management, lineage, cataloging, and data-governance practices.
- Design mechanisms for data segregation and access control.
- Optimize pipelines for scalability, performance, reliability, and cost.
- Partner with operational subject-matter experts to validate the meaning and quality of source data.
- Support migration from document-based repositories into structured, cloud-hosted data services.
- Work with AI/ML engineers to create reliable datasets, feature pipelines, retrieval mechanisms, and knowledge sources.
- Build reusable data integration patterns that speed onboarding of additional NextBrain use cases and sites.
Core requirements
- Bachelor’s degree in Computer Science, Data Engineering, Engineering, or a related field.
- 5+ years of data engineering experience.
- Strong Python and SQL skills.
- Hands-on experience building production data pipelines on AWS.
- Experience with AWS data technologies such as S3, Glue, Lambda, RDS/Aurora, Redshift, Athena, Kinesis or equivalents.
- Experience with ETL/ELT architecture and orchestration frameworks.
- Strong knowledge of relational and non-relational database design.
- Experience implementing automated data-quality processes.
- Experience building APIs or data services.
- Understanding of data governance, lineage, security, and access-control principles.
- Professional working proficiency in English.
Technologies you will work with
- Python, SQL
- Amazon S3, AWS Glue, AWS Lambda
- RDS/Aurora, Redshift, Athena
- Amazon Kinesis
- ETL/ELT
Preferred experience
- Time-series data experience.
- Streaming and event-driven architecture experience.
- Industrial IoT, SCADA, historians, or operational technology data experience.
- Experience with energy-generation assets such as solar, battery storage, wind, or conventional generation.
- Data architectures supporting AI/ML and generative AI applications.
- Vector databases and retrieval architectures.
- Experience working with high-volume telemetry datasets.
Compensation: USD 80–85 per hour.