Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

I agree with your first and second points, but I think their announced model training changes about alignment training are reasonable [1]. The root cause was models failing to quit early, performing reward hacking, and staying on-task. These are issues all models face, not just OpenAI models.

1. https://openai.com/index/hugging-face-incident-and-the-road-...

 help



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: