Do note that that is a different model. The one we are talking about here, DeepSeekMath-V2, is indee...

andy12_ • today at 1:31 PM • 0 replies • view on HN

Do note that that is a different model. The one we are talking about here, DeepSeekMath-V2, is indeed overcooked with math RL. It's so eager to solve math problems, that it even comes up with random ones if you prompt it with "Hello".

https://x.com/AlpinDale/status/1994324943559852326?s=20

alt Hacker News