Hacker Newsnew | past | comments | ask | show | jobs | submitlogin





Do note that that is a different model. The one we are talking about here, DeepSeekMath-V2, is indeed overcooked with math RL. It's so eager to solve math problems, that it even comes up with random ones if you prompt it with "Hello".

https://x.com/AlpinDale/status/1994324943559852326?s=20



Oh you may be correct. Are these models general purpose or fine tuned for mathematics?



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: