As the Research Engineer, you will design and run training experiments that isolate the impact of our datasets on model behavior. Through controlled SFT and RL - base post-training experiments, you will measure how different data sources, structures, and selection strategies affect capability, generalization, and alignment. You will own the full experiment loop: formulate the hypothesis, build the pipeline, run the model, analyze the results, identify confounders, and determine what should happen next. Working with partner labs and internal teams, you will turn our datasets into clear, defensible evidence: this data -> this improvement -> under these conditions. This is empirical, high-leverage work for someone who likes building quickly and extracting signal from noisy results.