Senior CE Data Engineer
Listed on 2026-08-31
-
IT/Tech
Data Engineering
Organization Overview
At Lilly, the work is demanding because patients are waiting. We unite caring with discovery to help make life better for people around the world, knowing that every decision, every detail, and every day matters. Headquartered in Indianapolis, Indiana, our over 50,000 employees around the globe take on complex challenges to discover and deliver life-changing medicines, strengthen how health is understood and managed, and support the communities we serve.
This is hard, urgent, selfless work—and it’s work worth doing. If you’re driven by purpose and ready to bring your best to work that truly matters for patients, we invite you to join us.
The CE Data Engineer builds and operationalizes the CE Pharmacy data platform on Databricks — implementing the medallion pipelines, data contracts, identity crosswalk logic, and isolation controls that the CE Data Architect designs. This is a hands‑on engineering role: writing production pipelines, evaluating and applying Databricks capabilities and integration patterns, and holding the line on data quality, testing, and monitoring so the platform is reliable, auditable, and built to scale.
Works in close, day‑to‑day partnership with the CE Data Architect and the Associate Director of CE Data Engineering — turning architectural intent into running code — and with Analytics Engineers and the Data Hub Product Owner to ship governed data products.
- Design and build scalable, efficient Databricks pipelines implementing the CE Architect's canonical data models across the medallion (bronze → silver → gold), scoped entirely within the CE trust boundary.
- Evaluate and apply Databricks capabilities and integration patterns — Unity Catalog, Delta Lake, Databricks Workflows, serverless compute, Lakebase, ingestion connectors — selecting the right tool for each pipeline given performance, cost, and scalability constraints.
- Implement the identity crosswalk and PHI classification logic defined by the CE Architect (Reltio MDM, Auth0/Passport, Datavant tokens, Scriptly, Genesys, Prescryptive, Transcend), safeguarding merge/split integrity in code.
- Build to ODCS data contracts — implement schema, quality, freshness/SLA, and lineage requirements as enforced pipeline logic, not documentation.
- Implement row/column-level security, masking, and tokenization boundaries so PHI isolation is enforced at the platform layer, in partnership with the Policy-as-Code Engineer's OPA/Rego policies.
- Own end-to-end pipeline development lifecycle — requirements to prototyping to production deployment and maintenance — for CE data products, in partnership with Analytics Engineers and the Data Hub Product Owner.
- Build and maintain CI/CD pipelines (Github Actions [CI], Git-based promotion dev → test → prod) for CE data, contract, policy, and agent artifacts.
- Automate data ingestion and product creation to reduce manual pipeline maintenance and onboarding time for new CE data sources.
- Partner with the CE Data Architect on reference architecture and patterns, providing implementation feedback that keeps designs buildable and performant at scale.
- Build the data pipelines that lets CE Skills and agents consume governed data — ensuring PHI classification and consent travel with the data into agentic consumption paths.
Bachelor's degree in Computer Science, Engineering, Information Technology, or similar degree. Experience in data engineering, with a focus on building production data pipelines and data products. Advanced SQL and Python; hands‑on Databricks / Unity Catalog fluency (notebooks, Delta Lake, catalog/schema/grants, Workflows). Qualified applicants must be authorized to work in the United States on a full‑time basis. Lilly will not provide support for or sponsor work authorization or visas for this role, including but not limited to F-1 CPT, F-1 OPT, F-1 STEM OPT, J-1, H-1B, TN, O-1, E-3, H-1B1, or L-1.
AdditionalSkills / Preferences
- Experience defining and executing data ingestion pipelines at enterprise scale.
- Working knowledge of data governance, classification, and access control (RBAC/ABAC, row- and column-level security).
- Profi…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).