https://x.com/deepseek_ai/status/1995452646459858977
Boom
Do note that that is a different model. The one we are talking about here, DeepSeekMath-V2, is indeed overcooked with math RL. It's so eager to solve math problems, that it even comes up with random ones if you prompt it with "Hello".
https://x.com/AlpinDale/status/1994324943559852326?s=20
That's a different model: https://huggingface.co/deepseek-ai/DeepSeek-V3.2-Speciale
Oh you may be correct. Are these models general purpose or fine tuned for mathematics?
Do note that that is a different model. The one we are talking about here, DeepSeekMath-V2, is indeed overcooked with math RL. It's so eager to solve math problems, that it even comes up with random ones if you prompt it with "Hello".
https://x.com/AlpinDale/status/1994324943559852326?s=20