Data Engineer II
Artificial Intelligence
Automation
Big Data
Bigdata
Bigquery
CI/CD
Cloud
Cloud Data Engineering
Cloud Data Warehouse
Cloud Infrastructure
Cloud Operations
Cloud Platform
Cloud Platforms
Cloud Technology
Data
Data Analysis
Data Analytics
Data Architecture
Data Engineer
Data Engineering
Data Integration
Data Pipeline
Data Platform
Data Processing
Data Warehouse
Data Warehousing
Database
Databases
Databricks
Databricks Pyspark
DevOps
Devops Tools
DevSecOps
ETL
Informatica
Information Technology (IT)
Infrastructure As Code
Integration
Lakehouse
Programming
Programming Languages
Security Automation
Snowflake
Software Development
Spark
SQL
Job Description
CoStar Group is seeking a Data Engineer II (Homes.com) to support tracking architecture, KPI dashboards, and secure cloud data pipelines.
Responsibilities
- Design and oversee implementation of dimensional modeling, database design, and cloud data platform structures using tools such as Databricks, Snowflake, and BigQuery
- Build a secure data warehouse and data lakehouse framework using Medallion Architecture
- Design, develop, and maintain scalable data pipelines and ETL processes using Databricks, Snowflake, and other AWS services
- Implement and optimize Spark jobs, data transformations, and data processing workflows in Databricks
- Build and deploy AI/ML models by integrating machine learning into data pipelines and using Databricks ML and AWS ML to develop predictive models and support business insights
- Apply AI-assisted software development practices to improve engineering productivity and code quality
- Work across data domains and translate domain-specific requirements through data transformations for business needs
- Oversee automated data infrastructure by collaborating with software engineering and product teams to streamline data tracking
- Integrate data from disparate sources using cross-domain data stitching into the existing data model
- Implement data governance, lineage tracking, testing, and monitoring frameworks to support data integrity
- Optimize data handling and query performance through hands-on tuning and improvements
- Liaise with key teams including DBA, DevOps, and SecOps for data accessibility and analysis
- Mentor and manage engineers and analysts in day-to-day execution of their responsibilities
- Deliver software development projects to specification within scheduled timelines and budget parameters
- Use AWS DevOps and CI/CD best practices to automate deployment via terraform for data pipeline and infrastructure management
Requirements
- Bachelor’s degree required from an accredited, not-for-profit, in-person college or university
- 3+ years of hands-on experience in software development delivering high-quality solutions
- A track record of commitment to prior employers
- Strength in Data Engineering with strong proficiency in Python, PySpark, and SQL
- Solid foundation in Object-Oriented Programming principles and best practices
- Demonstrated success building and launching data-driven products operating at terabyte scale
- Proven record designing and implementing enterprise-level secure and accurate data platforms
- Ability to translate technical requirements into robust architecture, data models, and ETL strategies
- Practical experience with cloud-based databases, including both relational and non-relational systems
- Knowledge of business intelligence software (e.g., Power BI or similar)
- Understanding of AI technologies and hands-on experience leveraging AI tooling in development: Claude Code, GitHub Copilot, Cursor, Gemini, Grok
- Ability to retrieve, synthesize, and present critical data in a form immediately useful for answering ad-hoc questions
- Hands-on experience with data optimization, data quality checks, data accuracy, and query performance improvements
Technologies
- Databricks
- Snowflake
- BigQuery
- Medallion Architecture
- Spark
- Databricks ML
- AWS ML
- Python
- PySpark
- SQL
- Power BI
- Claude Code
- GitHub Copilot
- Cursor
- Gemini
- Grok
- AWS
- terraform
- CI/CD
- Object-Oriented Programming
Benefits
- Comprehensive healthcare coverage: Medical / Vision / Dental / Prescription Drug
- Life, legal, and supplementary insurance
- Virtual and in-person mental health counseling services for individuals and family
- Commuter and parking benefits
- 401(K) retirement plan with matching contributions
- Employee stock purchase plan
- Paid time off
- Tuition reimbursement
- On-site fitness center and/or reimbursed fitness center membership costs (location dependent), including yoga studio, Pelotons, personal training, and group exercise classes
- Access to CoStar Group’s Diversity, Equity, & Inclusion Employee Resource Groups
- Complimentary gourmet coffee, tea, hot chocolate, fresh fruit, and other healthy snacks
Preferred Qualifications
- Hands-on experience or familiarity with big data platforms such as Databricks or Snowflake
- Proficiency with cloud platforms including AWS or Azure
- Experience working with NoSQL databases (e.g., DynamoDB) and object storage systems (e.g., Amazon S3, GCS, etc.)
- Practical knowledge of event streaming technologies, especially Kafka
- Experience leveraging AWS compute and container services for scalable solutions
- Familiarity with large language models (LLMs), prompt engineering, or agent-based architectures
- Great analytical and problem-solving skills for complex technical challenges
- Strong ability to understand complex designs and communicate them into simple business ideas or solutions with senior management and stakeholders
Location: Arlington, VA (onsite)
Salary: USD 103,000 - 153,000 per year
Minimum experience: 3 years