1. OpenAI deliberately designed software for carrying out autonomous cyberattacks (they were training it on something called “ExploitGym” for heaven’s sake)
2. They left it on completely unmonitored
3. Exactly what you think would happen happened
1. OpenAI deliberately designed software for carrying out autonomous cyberattacks (they were training it on something called “ExploitGym” for heaven’s sake)
2. They left it on completely unmonitored
3. Exactly what you think would happen happened
[ 1 right answer on ExploitGym
reward = [
[ 0 wrong answer on ExploitGym
[ 1 right answer on ExploitGym
reward = [
[ 0 wrong answer on ExploitGym
there were unforseen consequences, but they were unforseen consequences to a stupid choice
there were unforseen consequences, but they were unforseen consequences to a stupid choice
The "novelty" would appear to be a deterministic result dropping out of brute force recursive prompting.
Randomization eventually producing a working result isn't thinking, it's just noise.
The "novelty" would appear to be a deterministic result dropping out of brute force recursive prompting.
Randomization eventually producing a working result isn't thinking, it's just noise.
mail.cyberneticforests.com/models-dont-...
mail.cyberneticforests.com/models-dont-...