Data Engineer II
Artificial Intelligence
Automation
Big Data
Bigdata
Bigquery
CI/CD
Cloud
Cloud Data Engineering
Cloud Data Warehouse
Cloud Infrastructure
Cloud Operations
Cloud Platform
Cloud Platforms
Cloud Technology
Data
Data Analysis
Data Analytics
Data Architecture
Data Engineer
Data Engineering
Data Integration
Data Pipeline
Data Platform
Data Processing
Data Warehouse
Data Warehousing
Database
Databases
Databricks
Databricks Pyspark
DevOps
Devops Tools
DevSecOps
ETL
Informatica
Information Technology (IT)
Infrastructure As Code
Integration
Lakehouse
Programming
Programming Languages
Security Automation
Snowflake
Software Development
Spark
SQL
Job Description
As a Data Engineer II, you will support Homes.com by helping shape sitewide tracking architecture and ensuring data flows into KPI dashboards that inform day-to-day insights into consumer behavior. CoStar Group is focused on secure cloud data platforms, scalable pipelines, and practical analytics, with a role that also brings AI/ML model integration into the data lifecycle.
This onsite position is based in Arlington, VA. The salary range is USD 103,000 - 153,000 per year, and the role requires 3+ years of relevant hands-on experience. A bachelor’s degree from an accredited, not-for-profit, in-person college or university is required.
What you’ll do
- Design and oversee implementation of dimensional modeling, database design, and cloud data platform structures, including Databricks, Snowflake, and BigQuery
- Build a secure data platform framework for data warehouses and data lakehouses using Medallion Architecture
- Design, develop, and maintain scalable data pipelines and ETL processes using Databricks, Snowflake, and other AWS services
- Implement and optimize Spark jobs, data transformations, and data processing workflows in Databricks
- Build and deploy AI/ML models by integrating machine learning into data pipelines, leveraging Databricks ML and AWS ML to develop predictive models and support business insights
- Apply AI-assisted software development practices to improve engineering productivity and code quality
- Translate knowledge of multiple data domains into transformations that meet changing business needs
- Oversee automated data infrastructure and partner with software engineering and product teams to streamline product data tracking
- Integrate data from disparate sources using cross-domain data stitching to align with existing data models
- Implement data governance, lineage tracking, testing, and monitoring frameworks to help ensure data integrity
- Perform hands-on data optimization and query performance improvements
- Work with key teams (DBA, DevOps, SecOps) to support data accessibility and analysis
- Mentor and manage engineers and analysts in executing their normal responsibilities
- Deliver software development projects to specification, within scheduled timelines and budget parameters
- Use AWS DevOps and CI/CD best practices to automate deployments using terraform, and manage data pipelines and infrastructure
Skills and experience
- Bachelor’s degree required from an accredited, not-for-profit, in-person college/university
- Commitment track record with prior employers
- 3+ years of hands-on experience in software development delivering high-quality solutions
- Strong data engineering foundation with Python, PySpark, and SQL
- Solid understanding of Object-Oriented Programming principles and best practices
- Proven success building and launching data-driven products operating at terabyte scale
- Demonstrated experience designing and implementing enterprise-level secure and accurate data platforms
- Ability to translate technical requirements into robust architecture, data models, and ETL strategies
- Practical experience with cloud-based databases for both relational and non-relational systems
- Knowledge of BI software such as Power BI or similar tools
- Understanding of AI technologies and hands-on experience leveraging AI tooling in development, including Claude Code, GitHub Copilot, Cursor, Gemini, and Grok
- Ability to retrieve, synthesize, and present critical data in structures useful for answering ad-hoc questions
- Hands-on experience with optimization, data quality checks, data accuracy, and query performance improvement
Technologies you’ll work with
- Databricks, Snowflake, BigQuery
- Medallion Architecture
- Spark, Databricks ML, AWS ML
- Python, PySpark, SQL
- Power BI
- Claude Code, GitHub Copilot, Cursor, Gemini, Grok
- AWS, terraform, CI/CD
- Object-Oriented Programming
Benefits
- Comprehensive healthcare coverage: Medical / Vision / Dental / Prescription Drug
- Life, legal, and supplementary insurance
- Virtual and in person mental health counseling services for individuals and family
- Commuter and parking benefits
- 401(K) retirement plan with matching contributions
- Employee stock purchase plan
- Paid time off
- Tuition reimbursement
- On-site fitness center and/or reimbursed fitness center membership costs (location dependent), with yoga studio, Pelotons, personal training, group exercise classes
- Access to CoStar Group’s Diversity, Equity, & Inclusion Employee Resource Groups
- Complimentary gourmet coffee, tea, hot chocolate, fresh fruit, and other healthy snacks