You will be responsible for support and development of software to handle data curation, coordination, validation, distribution and visualisation of agriculture, aquaculture and biodiversity data. You will work with large scientific communities to define and implement metadata standards, develop software to validate and improve data descriptions and develop, build new project specific portals, and extend existing web data portals to meet the specific needs of each consortium.
The data portals will include Application Programming Interfaces (APIs) and bespoke data visualisation and presentation solutions. You will also contribute to our growing “question in → insight out” initiative, exposing consortium data through MCP servers and agent-based interfaces so that researchers can ask scientific questions in natural language and get back analysis, not just download links. You will also support efforts for the development of standardised containerised workflows and cloud platform integration.
Reporting to our Genome Analysis Team Leader, you’ll be part of a high performing team enjoying many opportunities to engage with data generators, project users, and collaborators. As part of this you will provide valuable guidance and support on the utilisation of the team’s software and support users to provide rich metadata descriptions.
You will work with a range of data archives at EMBL-EBI, including the ENA and BioSamples, to support each project and community in sharing and gaining access to well described, high quality sample and genomic data.