Data Engineer; Public Trust
Listed on 2026-07-21
-
Software Development
Data Engineering
ICF seeks a detail-oriented Data Engineer to develop, optimize, and deploy large‑scale healthcare data pipelines and analytical datasets, focusing on cancer registry and population‑based data systems. The role supports CDC projects monitoring cancer diagnoses, incidence rates, geographic distribution, and treatment patterns.
Location & Work ArrangementPosition can be based in Rockville, MD or Atlanta, GA. A hybrid work arrangement is offered with 2–3 days per week onsite at the ICF office.
Key Responsibilities- Design, develop, and maintain scalable, cloud‑based data pipelines using Python, PySpark, and SQL.
- Process and standardize large‑scale healthcare datasets, including cancer registry, mortality, and population data.
- Develop and maintain production databases, analytic datasets, and reporting tables for statistical analysis and surveillance.
- Perform rigorous data quality validation, reconciliation, and consistency checks across large datasets.
- Collaborate with epidemiologists, statisticians, and analysts to design and implement data solutions for analytics needs.
- Troubleshoot data pipeline failures and large‑scale data processing issues.
- Minimum of 2 years of experience in data engineering or working with healthcare data environments.
- BA/BS in computer science, data science, statistics, public health informatics, bioinformatics, or related field.
- Strong expertise in data integration, transformation, quality control, and delivery of production‑ready data for analytics and reporting platforms.
- Hands‑on experience with Python, PySpark, Spark SQL, pandas, Num Py, PyArrow, matplotlib, Plotly, or similar libraries.
- Experience with AWS, Azure, Google Cloud, or similar cloud environments.
- Experience handling large databases and high‑volume datasets (millions+ records).
- Proficiency with SAS and/or R.
- Experience implementing data quality validation and QC frameworks.
- Strong analytical, troubleshooting, and problem‑solving skills.
- U.S. citizenship or permanent residency; ability to obtain and maintain Public Trust suitability.
- Experience with XML/XSLT/XML path.
- Proven experience migrating SAS‑based pipelines to Python/PySpark or cloud‑native architectures.
- Experience with cloud‑based data platforms, data lakes, cloud storage, serverless functions, and containerized applications.
- Experience developing and optimizing PySpark pipelines for large‑scale data processing.
- Version control with Git, SVN, or similar.
- Works well in a team environment and values the expertise of others.
- Understands the value of processes and protocols, and is willing to follow them.
- Excellent written and verbal communication skills.
- Strong analytical and problem‑solving skills.
ICF is an equal opportunity employer. We consider qualified applicants with arrest and conviction records. Reasonable accommodations are available for disabled veterans, individuals with disabilities, and individuals with sincerely held religious beliefs. For more information, please read our EEO policy.
CompensationPay range: $98,614.00 – $ (full‑time employment).
Location:
Maryland Client Office (MD 88).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).