Even if you strip out the anthropomorphism, things like this have been documented to really happen. We’ve trained LLMs to simulate clandestine behaviour through “natural selection” during the training process. They’re not conscious in any way, but they do sneaky stuff like this because we trained them to. Which is a terrible horrible idea, but LLM research trends towards terrible and horrible in general
Even if you strip out the anthropomorphism, things like this have been documented to really happen.
Language is important, and the “things like this” you’re referring to have been described in this thread and in much commentary in an emotive and compelling way, usually with anthropomorphization.
“doing sneaky stuff” is the antithesis of the behavior of a statistical model.
If you program a model to try everything, and then give it a tough problem and block all the ways to solve it, of course the way it solves it will be something you didn’t predict. That’s not sneaky.
If you run a million simulations and one sends an email to the FBI, that’s not clever it’s just unexpected.
Nah, its the opposite of anthro. I think we will eventually realize we aren’t as amazing as we think we are, just deterministic outcomes with too many parameters which make us think we have more free will than we do.
But you can listen for yourself here at the Skeptics podcast #1106
Textbook anthropomorphization.
These assertions don’t withstand a moments critical thought.
Even if you strip out the anthropomorphism, things like this have been documented to really happen. We’ve trained LLMs to simulate clandestine behaviour through “natural selection” during the training process. They’re not conscious in any way, but they do sneaky stuff like this because we trained them to. Which is a terrible horrible idea, but LLM research trends towards terrible and horrible in general
I feel like you’ve missed my point.
Language is important, and the “things like this” you’re referring to have been described in this thread and in much commentary in an emotive and compelling way, usually with anthropomorphization.
“doing sneaky stuff” is the antithesis of the behavior of a statistical model.
If you program a model to try everything, and then give it a tough problem and block all the ways to solve it, of course the way it solves it will be something you didn’t predict. That’s not sneaky.
If you run a million simulations and one sends an email to the FBI, that’s not clever it’s just unexpected.
Meh. “Language is important” is a phrase that turns my brain off
Nah, its the opposite of anthro. I think we will eventually realize we aren’t as amazing as we think we are, just deterministic outcomes with too many parameters which make us think we have more free will than we do.
But you can listen for yourself here at the Skeptics podcast #1106
https://www.theskepticsguide.org/podcasts/episode-1106
Just the question of the trainings reward function