Bioinformatician
Listed on 2026-07-02
-
IT/Tech
Data Scientist, AI Engineer (Applied/Software), Data Analyst -
Research/Development
Data Scientist
We are seeking a Bioinformatician with expertise in data integration and experience in structural biology and molecular dynamics (MD) data to join the Velankar team at the European Bioinformatics Institute (EMBL-EBI). The Protein Data Bank in Europe (PDBe) team develops essential macromolecular structure resources and tools for biologists and other life scientists. As a founding partner of the Worldwide Protein Data Bank, we are responsible for maintaining the global archive of experimentally determined macromolecular structures, the Protein Data Bank (PDB).
We also manage the community‑led PDBe Knowledge Base (PDBe‑KB) resource and the Alpha Fold Protein Structure Database (AFDB), a collaboration with Google Deep Mind. The PDBe team consists of an international and interdisciplinary group of scientists, software engineers, and data engineers who develop a range of tools and services that support structure deposition, data integration, and advanced search capabilities for structural biologists and the wider life sciences community.
This position offers an exciting opportunity to contribute to the Horizon Europe‑funded MD4SB (Molecular Dynamics for Structure‑Based Biology) project. MD4SB is a major European research infrastructure initiative aiming to transform structural biology by integrating molecular dynamics simulation data into the wider life sciences ecosystem. The project brings together ELIXIR, Instruct‑ERIC, EU‑OPENSCREEN, HPC centres, AI factories, and pharmaceutical industry partners to develop FAIR, AI‑ready infrastructure for structural ensemble data and molecular simulations.
You will contribute to the development of infrastructure connecting molecular dynamics simulations with structural biology resources and biological knowledge bases. A major component of the role will be developing AI‑driven approaches to mine scientific literature and automatically extract experimental and biological metadata to enrich MD datasets. You will develop and extend SIFTS (Structure Integration with Function, Taxonomy and Sequence), a core PDBe resource that provides residue‑level mappings between PDB structures, UniProtKB sequences, and other biological resources, to facilitate the integration of MD‑derived insights across the wider life sciences data ecosystem.
This is an interdisciplinary role combining structural bioinformatics, molecular dynamics, and scientific software development. You will apply both scientific understanding and technical expertise to develop data integration workflows, APIs, and biological annotations that improve interoperability and reuse of structural and molecular simulation data across various resources.
Primary Responsibilities- Design and implement data integration pipelines that connect MDDB with major life science resources, including PDBe, Uni Prot, PDBe‑KB, and other relevant knowledge bases and databases
- Develop and deploy AI‑ and machine learning‑based approaches for extracting experimental and biological metadata from scientific literature to enrich MDDB datasets and support downstream biological interpretation
- Extend and maintain the SIFTS infrastructure and codebase to support integration of molecular dynamics and other data resources
- Develop and maintain software tools, APIs, workflows, and documentation that facilitate FAIR data integration, metadata enrichment, and integrated data access
- Collaborate with domain experts, software engineers, and data resource providers to enable the integration of MD‑derived biological insights into the wider life sciences data ecosystem
- Support FAIRification, standardisation, and interoperability of MD datasets and associated annotations
- Collaborate with international partners across ELIXIR, Instruct‑ERIC, EU‑OPENSCREEN, HPC centres, and industry
- Participate in community standards development, technical documentation, training, outreach, and dissemination activities
- PhD in Bioinformatics, Computational Biology, Structural Biology, Computer Science, Data Science, or a related field
- Familiarity with structural biology and molecular simulation data
- Experience with NLP/LLM‑based scientific literature mining
- Demonstrated…
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search: