You have a deep understanding of biological data and a strong core in sequence informatics (alignment, annotation, sequence analysis), with the judgment to know when standard tools are misleading or not right for the job. When a tool or method doesn’t exist, your instinct is to build one - and the tools and methods you’ve built have been applied to real data at scale.
You’ve curated data that went on to train real models or feed large-scale analyses. You’re a capable enough engineer to make your methods real, but you treat engineering as the means, not the mission.
You like to move fast, to ship, learn, and iterate at pace, without cutting corners on rigour.
You are excited by the scale of the biology: genetic diversity that exists in no public database, and the chance to invent the methods that make it usable for the first time.
You’ll likely have a PhD in a relevant field, an MSc with 1-3 years of relevant experience, or a BSc with 3 or more.
Low ego, collaborative instincts, and a startup mentality. You’re comfortable with ambiguity and happy to wear multiple hats in a team where everyone contributes towards the company’s success.