In the past, community discussions on improving reasoning abilities often focused on optimizing RL algorithms or constructing verifiable data in domains like Math and Code. In the M2 project, we conducted more “general” explorations. As a member of the Reasoning team, I’d like to share some of our findings and thoughts on data — what makes good reasoning data.