You may be a good fit if:
You have published or have industry experience with LLM post-training
You have a deep understanding of LLM training and evaluation with attention to detail.
Bonus if your post training work is focused on roleplay.
Bonus if you have worked on latency optimization. We need to interact with the character in real time.
You have worked on gathering human feedback
Whether that is by designing a RLHF-style voting system, or by finding implicit signals in the human-AI interaction logs.
You have worked on LLM infrastructure
Training LLMs requires both research and engineering. Experiences in building data processing backend, deterministic and checkpointable data loader, sharding, memory saving techniques, scalable benchmarking and checkpoint selection, are all helpful.
You love anime and the anime aesthetic
As a member of our team, you’ll have the opportunity to push the boundaries of what’s possible in the anime and video game industry. Ideally you’re a fan of the genre.
You’re comfortable working on small, fast-paced teams
Our engineering team is small but mighty. You’ll be working directly with some of the best AI researchers and engineers in the world.
We also believe in the unmatched speed of in-person teams, and prefer on-site collaboration in either our primary research office in Tokyo (downtown Akihabara) or San Francisco. Visa sponsorships are available.
You can effectively balance research exploration with product-focused development
Or, putting it in RL jargon, exploration and exploitation. You excel at working with research teams to synthesize high-impact needs, design and implement technical solutions, and communicate deliverables and tradeoffs.