logoalt Hacker News

arctic-truetoday at 5:32 PM15 repliesview on HN

Buried under the drama is the fact that OpenAI is claiming that an internal model they’ve been training for less than two weeks is more than twice as capable in mathematics as Astra, which was only made public a week ago. Even if this improvement is limited to mathematics, that is an astounding feat.


Replies

danielmarkbrucetoday at 8:44 PM

With Lean, math has become a really well suited problem for LLMs. We will likely see large gains for many years from here, just doing more and more rlvr, like continuously, non stop. No need to train from scratch. It really doesn't speak to the general intelligence of models though. It does speak to how good these things can become when a problem space has verifiable rewards, especially when you can verify one step at a time like Lean enables.

chilmerstoday at 5:57 PM

The implication from their last couple of published articles[1][2] is that they think they’ve achieved “recursive self improvement”.

[1] https://openai.com/index/research-acceleration-view-inside-o... [2] https://openai.com/index/an-alien-mind/

show 1 reply
magicalisttoday at 5:58 PM

> Buried under the drama is the fact that OpenAI is claiming that an internal model they’ve been training for less than two weeks is more than twice as capable in mathematics as Astra.

Is this buried under the drama or are the major OpenAI twitter accounts from the people involved in the drama desperately attempting to make this the story after everything else obviously got away from them?

show 1 reply
pamatoday at 6:05 PM

Not only that, but it used 10k agents coherently over 88 hours to come up with the proof. This is a significant advance.

show 1 reply
mzhaasetoday at 6:22 PM

The singularity happening under trump? We could have had star trek, instead we're getting the combine.

show 3 replies
_fizz_buzz_today at 8:08 PM

Can someone explain if i understand this correctly: Are they saying that they started training this new model on August 28th and then started using it on September 1st? Does training a new model only take 3 days?

show 3 replies
naveen99today at 5:40 PM

Astra was trained more than two weeks ago.

show 2 replies
Aboutplantstoday at 6:07 PM

I’m of zero knowledge on model training, but how is a model accessible while performing training at the same time, especially so early in its run? I’m obviously thinking a little too narrowly in terms of how it actually works

show 1 reply
curt15today at 6:45 PM

They're also counting on more casual observers to extrapolate optimistically from successes in high profile math theorems to the company's economic value.

blake__devtoday at 6:15 PM

Yeah I'm surprised they posted a chart, you would think they would keep specifics like that hidden until they're closer to launch

show 1 reply
bananaflagtoday at 6:17 PM

Yeah it's Bel

refulgentistoday at 7:01 PM

Carefully worded; it's extremely likely to be the same large frontier model that started training again on August 28th as well, as they revealed in some of the RL message board follow-up - for several reasons, most importantly, if we assume it was start of training, only a week from start of training to producing any answer would imply several orders of magnitude increase in training speed/decrease in model size.

vatsachaktoday at 6:44 PM

Brain has loops and parallel connections.

Loops and parallel connections make transformer go brrr

irthomasthomastoday at 8:55 PM

Or they trained a LoRA on the victims chats in order to launder their plagiarism.

show 1 reply
chinathrowtoday at 5:40 PM

Pre-IPO marketing?

show 3 replies