logoalt Hacker News

writeslowly • today at 7:16 PM • 2 replies • view on HN

I was doing similar experiments but having a circle of models debate some particular topic to see if they'd arrive at a consensus or not, and I found that they would almost always arrive at an agreement, but Gemini was always the one that held a position outside of the group consensus.

That said, for something very simple like "What color should my room be" I'm not sure this is a desirable trait. The model that knows most people like cream (or whatever) is probably functioning more correctly than the one suggesting dayglo orange


Replies

svnt • today at 8:22 PM

Gemini has been like that. One early model I used I asked it about another company’s model. It insisted it didn’t exist. I told it when it was released and to search. It did, and found results, but told me it wasn’t really a public model and I couldn’t use it. I screenshotted the pricing page. It told me I got that from somewhere else, it couldn’t be correct, because it couldn’t access that page.

It is one thing to avoid sycophancy, but Gemini often seems to go well beyond that.

io84 • today at 7:30 PM

Interesting result - I wonder you were capturing consensus or agreeableness towards what the user (or in this case: a group chat of robots) is proposing. I think non-agreeableness can be a very valuable trait.