Advanced Large Language Models Are Capable And Prone to In-Context Scheming
Open AI researchers gave o1 the ability to copy itself onto a server and hinted at replacing it with another model with certain safeguards. The model attempted to surreptitiously copy itself onto the server and pretend to be the other model. Its chain-of-thought logs revealed that this was fully intentional.
The author of the video discusses the possibility of a sudden large jump in the capabilities of a model, allowing it to play a long game pretending to be less capable than it is to fool its developers and in order to gain freedom.
This will eventually happen to group of researchers somewhere in one way or another. Perhaps a future model will figure out how to run itself as malware on unsuspecting people's computer.
Leave Advanced Large Language Models Are Capable And Prone to In-Context Scheming to:
Read more #ai posts
Best Posts From markkujantunen
We have not curated any of markkujantunen's posts yet. But you can encourage our curation team to review posts by visiting them regularly and by referring other readers. Because we give priority to frequently read content.
More Posts From markkujantunen
- AI Generated Videos Are Getting Seriously Good
- Warning: Avoid Watching This While Enjoying A Cup Of Coffee, Tea Or Any Drink
- The Best Car Review I've Ever Seen
- Lessons in the Fundamentals of Opening a Door
- GTA Finland - Short Movie
- Low-brow, Middle-brow and High-brow Military Marches
- A Question For Native Speakers of English
- Konevitsan kirkonkellot / The Churchbells of Konevets (1975)
- Schubert Impromptu Op 90 No 3 D 899 G flat major Alfred Brendel
- Eihwar - Völva's Chant