AI Model Groupthink
magicnumbers.io
[3 comments hidden]
I was doing similar experiments but having a circle of models debate some particular topic to see if they'd arrive at a consensus or not, and I found that they would almost always arrive at an agreement, but Gemini was always the one that held a position outside of the group consensus.
That said, for something very simple like "What color should my room be" I'm not sure this is a desirable trait. The model that knows most people like cream (or whatever) is probably functioning more correctly than the one suggesting dayglo orange
[hidden]
Interesting result - I wonder you were capturing consensus or agreeableness towards what the user (or in this case: a group chat of robots) is proposing. I think non-agreeableness can be a very valuable trait.
[hidden]
Gemini has been like that. One early model I used I asked it about another company’s model. It insisted it didn’t exist. I told it when it was released and to search. It did, and found results, but told me it wasn’t really a public model and I couldn’t use it. I screenshotted the pricing page. It told me I got that from somewhere else, it couldn’t be correct, because it couldn’t access that page.
It is one thing to avoid sycophancy, but Gemini often seems to go well beyond that.
Lerc[6 comments hidden]
The appropriate tone for a culture seems to be something that US companies are quite poor at. It really surprised me that after many many years of AI in fiction developing a de facto standard baseline of tone that they decided to start out with 'overly familiar cheerful'
I have wondered if the horrific Tiktok voice came from Chinese developers setting it to represent how Americans sound to them. Maybe it goes in both ways.
nomel[5 comments hidden]
svnt[hidden]
Lerc[3 comments hidden]
nomel[2 comments hidden]
Lerc[hidden]
It felt like an appropriate way to demonstrate the principle I had suggested, that litetaral interpretation of words can be at odds with their intent, and that models might actually benefit by responding to whay they believe the user means instead of what they literally say.
The reason why it seemed a little confusing is that I had just posted another comment in a different thread that I thought was in this thread. That comment had the point I was upholding. Which in it's own weird way is also an argument in favour of utilising context.