Senior Data Engineer
Richmond, Henrico County, Virginia, 23214, USA
Listed on 2026-08-26
-
Software Development
Data Engineering, SQL Developer
Senior Data Engineer
Job Description
CoStar Group (NASDAQ: CSGP) is a leading global provider of commercial and residential real estate information, analytics, and online marketplaces. Included in the S&P 500 Index, CoStar Group is on a mission to digitize the world’s real estate, empowering all people to discover properties, insights and connections that improve their businesses and lives.
We have been living and breathing the world of real estate information and online marketplaces for over 35 years, giving us the perspective to create truly unique and valuable offerings to our customers. We’ve continually refined, transformed and perfected our approach to our business, creating a language that has become standard in our industry, for our customers, and even our competitors.
We continue that effort today and are always working to improve and drive innovation. This is how we deliver for our customers, our employees, and investors. By equipping the brightest minds with the best resources available, we provide an invaluable edge in real estate.
We are seeking a Senior Data Engineer to join a new team building our next generation of data warehousing pipelines and ecosystem in the finance tech organization. Initially this role will focus on building a data migration and transformation process for external systems for use in our enterprise contracting system. This will include receiving data from multiple data sources of varying format, frequency and type.
You will create applications to process it in a repeatable, customizable and recurring state to simplify the conversion process of our financial data across systems and acquisitions both today and in the future. You will also help standardize our data environment across our billing systems to allow for greater visibility into patterns, errors and optimizations to our business, consisting of billions of dollars of payments a year.
This position is in Richmond, VA and has a work schedule of Monday through Thursday in office and Friday work from home.
Responsibilities- Design, build, and maintain scalable data pipelines and data platforms
- Develop and optimize ETL/ELT processes for large, complex datasets
- Collaborate with engineering, product, Dev Ops and DBA teams to deliver production systems
- Contribute to technical design decisions and mentor other engineers on the team
- Bachelor's degree required from an accredited, not-for-profit, in-person college/university
- Track record of commitment to prior employers
- 5+ years of data engineering experience, including experience setting technical direction on projects and influencing design across a team
- Hands‑on experience building batch and incremental ingestion pipelines from multiple heterogeneous external sources (REST/GraphQL APIs, flat-file and SFTP feeds, third‑party data vendors, and database replication/CDC)
- Demonstrated experience normalizing inconsistent third‑party schemas into standardized enterprise data models, including source‑to‑target mapping, deduplication, and entity resolution
- Advanced SQL and strong data modeling skills (dimensional/star schema, normalized, or Data Vault), plus proficiency in Python, Scala, Java, or C#/.NET
- Production experience with a distributed processing engine and a cloud data warehouse or lakehouse platform (e.g., Databricks/Spark, Snowflake, Big Query, Redshift, Synapse/Fabric), including pipeline orchestration tooling (e.g., Airflow, Dagster, Databricks Workflows, Azure Data Factory)
- Experience implementing data quality, validation, and reconciliation controls for externally sourced data, with version control and CI/CD in a major cloud environment (AWS, Azure, or GCP)
- Experience with agentic engineering with AI-assisted development tools (Claude Code or similar) to accelerate software delivery
- Production‑scale expertise with Databricks (Delta Lake, Unity Catalog, Delta Live Tables / declarative pipelines, Databricks SQL, Workflows) or an equivalent lakehouse platform, including open table formats (Delta Lake, Apache Iceberg, Apache Hudi)
- Experience with a modular transformation framework and tested, version‑controlled…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).