DataJobs.io
← Back to all jobs

Job Description

Vertage is building NextBrain, a cloud data foundation meant to securely and reliably power access to operational, engineering, analytical, and contextual information. As a Senior Data Engineer in the United States (onsite), you will design and deliver the core pipelines, storage architecture, data models, APIs, and automated data-quality controls on AWS.

This role focuses on making NextBrain data dependable for both applications and AI capabilities, with careful attention to governance, lineage, access control, and the performance and cost of production workloads.

What you will do

  • Design and implement the NextBrain cloud data architecture.
  • Build scalable ingestion, transformation, storage, and serving pipelines on AWS.
  • Create pipelines for operational, historical, engineering, application, and contextual data.
  • Design data models optimized for analytical applications and AI consumption.
  • Define storage patterns across relational, object, time-series, and analytical data stores.
  • Develop APIs and services that let NextBrain tools and agents retrieve data consistently.
  • Implement automated data-quality validation, reconciliation, and monitoring.
  • Establish metadata management, lineage, cataloging, and data-governance practices.
  • Design mechanisms for data segregation and access control.
  • Optimize pipelines for scalability, performance, reliability, and cost.
  • Partner with operational subject-matter experts to validate the meaning and quality of source data.
  • Support migration from document-based repositories into structured, cloud-hosted data services.
  • Work with AI/ML engineers to create reliable datasets, feature pipelines, retrieval mechanisms, and knowledge sources.
  • Build reusable data integration patterns that speed onboarding of additional NextBrain use cases and sites.

Core requirements

  • Bachelor’s degree in Computer Science, Data Engineering, Engineering, or a related field.
  • 5+ years of data engineering experience.
  • Strong Python and SQL skills.
  • Hands-on experience building production data pipelines on AWS.
  • Experience with AWS data technologies such as S3, Glue, Lambda, RDS/Aurora, Redshift, Athena, Kinesis or equivalents.
  • Experience with ETL/ELT architecture and orchestration frameworks.
  • Strong knowledge of relational and non-relational database design.
  • Experience implementing automated data-quality processes.
  • Experience building APIs or data services.
  • Understanding of data governance, lineage, security, and access-control principles.
  • Professional working proficiency in English.

Technologies you will work with

  • Python, SQL
  • Amazon S3, AWS Glue, AWS Lambda
  • RDS/Aurora, Redshift, Athena
  • Amazon Kinesis
  • ETL/ELT

Preferred experience

  • Time-series data experience.
  • Streaming and event-driven architecture experience.
  • Industrial IoT, SCADA, historians, or operational technology data experience.
  • Experience with energy-generation assets such as solar, battery storage, wind, or conventional generation.
  • Data architectures supporting AI/ML and generative AI applications.
  • Vector databases and retrieval architectures.
  • Experience working with high-volume telemetry datasets.

Compensation: USD 80–85 per hour.

Similar Jobs