logoalt Hacker News

peterBlue75 • today at 5:51 PM • 1 reply • view on HN

The thing to note here, besides the transparency and the fact that it’s actually a good model that also works well on coding and agentic tasks, is that it’s the first release by a team formed less than a year ago, with a strong focus on iteration velocity. There’s more to come.

disclaimer: I‘m part of the training team, happy to answer any questions


Replies

ducktective • today at 7:22 PM

- Is it possible to train only on math and logic materials and expect the model's response in math questions to be superior to general models with the same training/inference compute hardware?

- Are there non-LLM approaches to the above task with the goal of achieving a non-hallucinatory agent?