logoalt Hacker News

dpoloncsakyesterday at 1:45 PM1 replyview on HN

I will say, when working with the recent batch of frontier models, even the last batch, if you ask "Do you think we should xyz" it may sometimes push back for an alternate solution.

Now, I can't promise it's advice is worth taking, but they do seem to be taking strides at judging the relevancy of some needless additions.


Replies

spprashantyesterday at 1:52 PM

I agree with your observation. Although the framing of the question signals some of the response "Do you think.. " forces the model to check for the cost-benefits.

We can definitely mould AI agents to think more critically about these things, I dont know how effective it will be in the long term honestly.

show 1 reply