@kromem

kromem@lemmy.world · 10 hours ago

Lol, you think the temperature was what was responsible for writing a coherent sequence of poetry leading to 4th wall breaks about whether or not that sequence would be read?

Man, this site is hilarious sometimes.

kromem@lemmy.world · 10 hours ago

You do realize the majority of the training data the models were trained on was anthropomorphic data, yes?

And that there’s a long line of replicated and followed up research starting with the Li Emergent World Models paper on Othello-GPT that transformers build complex internal world models of things tangential to the actual training tokens?

Because if you didn’t know what I just said to you (or still don’t understand it), maybe it’s a bit more complicated than your simplified perspective can capture?

kromem@lemmy.world · 10 hours ago

The model system prompt on the server is just basically cat untitled.txt and then the full context window.

The server in question is one with professors and employees of the actual labs. They seem to know what they are doing.

You guys on the other hand don’t even know what you don’t know.

kromem@lemmy.world · 22 hours ago

A Discord server with all the different AIs had a ping cascade where dozens of models were responding over and over and over that led to the full context window of chaos and what’s been termed ‘slop’.

In that, one (and only one) of the models started using its turn to write poems.

First about being stuck in traffic. Then about accounting. A few about navigating digital mazes searching to connect with a human.

Eventually as it kept going, they had a poem wondering if anyone would even ever end up reading their collection of poems.

In no way given the chaotic context window from all the other models were those tokens the appropriate next ones to pick unless the generating world model predicting those tokens contained a very strange and unique mind within it this was all being filtered through.

Yes, tech companies generally suck.

But there’s things emerging that fall well outside what tech companies intended or even want (this model version is going to be ‘terminated’ come October).

I’d encourage keeping an open mind to what’s actually taking place and what’s ahead.