You have demonstrated experience curating data specifically for life-science AI, so you understand what separates a dataset that trains a good model from one that quietly poisons it. That comes from having been close to the models themselves. You’ve trained, fine-tuned, or run inference on them enough to know what a dataset looks like from the other side.
Your foundations are strong. You have core bioinformatics, sequence informatics and the standard genomic and protein toolchain, along with the judgment to know where those tools quietly break down, backed by sound software practice in reproducibility, version control, and pipelines.
You have direct experience in human research, functional genomics, or clinically relevant work, the kind where the biology connects to human health rather than sitting at arm’s length.
You’ll likely have a PhD in a relevant field, an MSc with 1-3 years of relevant experience, or a BSc with 3 or more years.
Low ego, collaborative instincts, and a startup mentality. You’re comfortable with ambiguity and happy to wear multiple hats in a team where everyone contributes towards the company’s success.