Awesome work! Wondering if you will extend the current work to the post training like RL/SFT/OPD and use them for the auto-research?
Awesome work! Wondering if you will extend the current work to the post training like RL/SFT/OPD and use them for the auto-research?