Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Yes, that does sound very similar. To my knowledge, isn’t that (effectively) how the latest DeepSeek breakthroughs were made? (i.e. by leveraging chatgpt outputs to provide feedback for training the likes of R1)


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: