Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

My understanding is that RLVR, synthetic data generation and a slew of other post-training techniques are what have driven many recent advances in models more so than manual data providers. The economics of that are for sure worse than just scaling pre-training but it is incorrect to think that test time inference scaling and manual data entry are the only ways in which models are advancing.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: