Data Engineer 5 (SEO/GEO, AdTech )
Job Description
Capital One is seeking a Data Engineer 5 for SEO/GEO in AdTech to support a major transformation through cloud-first data solutions and scalable marketing or messaging platforms. The role includes designing, developing, testing, implementing, and operating data pipelines and platforms, while mentoring other engineers and contributing to architectural decisions.
Location and Role Details
- Location: McLean, VA (onsite)
- Experience: 6+ years
- Salary: USD 229,900 - 262,400 per year
- Education: Bachelor’s Degree in Computer Science or a related quantitative field
What You’ll Do
- Collaborate with and across Agile teams to design, develop, test, implement, and support technical solutions using full-stack development tools and technologies
- Influence a team of developers, data analysts, and data scientists through experience spanning machine learning, distributed microservices, lakehouse architecture, and full-stack systems
- Build data pipelines and platforms using Python and Spark, along with open-source relational and NoSQL databases and cloud data warehousing platforms such as Databricks and Snowflake
- Collaborate with product managers and software engineers to deliver robust cloud-first data solutions that support powerful experiences for millions of Americans
- Independently design, build, and deliver cloud data solutions and applications with minimal support from supervisors or managers
- Architect and enforce common data engineering design patterns to improve code quality, maintainability, and reusability across platforms and pipelines
- Ensure scalability, resilience, and operational efficiency by designing pipelines and platforms that maintain robust performance as data volumes and business demands increase
- Serve as a data engineering ambassador by communicating technical concepts and data outcomes clearly to internal and external stakeholders
- Act as a force-multiplier by balancing hands-on engineering contribution and innovation with mentoring and skill development for peers and junior engineers
- Lead end-to-end, large-scale data initiatives, including critical architectural decisions and platform evaluations such as Snowflake versus Databricks based on technical and business requirements
- Stay current with data engineering and technology trends, experiment with new technologies, and participate in internal and external technology communities while mentoring members of the data community
Technologies
- Python, Spark, SQL
- Databricks, Snowflake
- Java, Scala
- EMR, Glue, Redshift
- Airflow, Dagster
- Monte Carlo, Splunk
- AWS, Microsoft Azure, Google Cloud
- NoSQL, MongoDB, Cassandra, DynamoDB
Required Qualifications
- Bachelor’s Degree or higher in Computer Science or a related quantitative field (Statistics, Economics, Operations Research, Analytics, Mathematics, Engineering)
- 6+ years of experience in application development (internship experience does not apply)
- 4+ years of experience in distributed data
- 4+ years of experience with SQL
- 4+ years of programming with at least one of: Python, Java, or Scala
- 4+ years of experience designing and developing data pipelines
- 2+ years of experience in data modeling and end-to-end data solutions using both relational and non-relational database systems
Team and Focus
The Marketing and Messaging team is responsible for delivering hyper-personalized messages and experiences that attract prospects, delight customers, and drive increasing business value. The team builds scalable platforms that deliver omnichannel messages in owned and paid AdTech channels.
Preferred Qualifications
- Master’s Degree in Computer Science or a related field
- 8+ years of experience in data engineering
- 4+ years of data modeling experience
- 9+ years of experience in application development with demonstrated proficiency in Python, SQL, Scala, or Java
- 5+ years of hands-on experience designing, deploying, and operating data workloads in at least one public cloud environment (AWS, Microsoft Azure, or Google Cloud)
- 5+ years of experience building or supporting distributed data or compute workloads using tools such as EMR, Spark, Glue, or Databricks
- 5+ years of experience designing, implementing, and operating real-time or streaming data pipelines
- 3+ years of experience working with data observability (e.g., Monte Carlo, Splunk) or data orchestration tools (e.g., Airflow, Dagster)
- 5+ years of experience working with unstructured or semistructured data using NoSQL databases (e.g., MongoDB, Cassandra, DynamoDB)
- 5+ years of experience designing and supporting data warehousing solutions (e.g., Snowflake, Redshift)
- 3+ years of experience working in an Agile development environment
- 3+ years of experience developing user-centric reusable data products
Benefits
- Comprehensive, competitive, and inclusive set of health, financial, and other benefits supporting total well-being
- Performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI)
Additional Information
- This role is also eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI)
- Capital One is expected to accept applications for a minimum of 5 business days
- No agencies please
- Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination
- Capital One promotes a drug-free workplace
- Capital One will consider for employment qualified applicants with a criminal history in a manner consistent with applicable laws
- If you require an accommodation, contact Capital One Recruiting at 1-800-304-9102 or via email at [email protected]
- For technical support or questions about Capital One’s recruiting process, send an email to [email protected]