The Hard Problem of Controlling Powerful AI Systems - Computerphile
62K views · Dec 4, 2025 · Education
Comments · 333
@davidswinstead · 9 months ago
You ask an agent to find a machine's IP address, go for coffee, and when you come back the agent has SSHed in and upgraded the operating system..... my guy you are WAY TOO CHILL about this haha
202
@ricoreyes6044 · 9 months ago
It's like the Simpsons episode where Homer puts the drinking bird at the computer - "Why did I leave you in charge?!"
81
@BobbyKingOfTheSeas · 9 months ago
Be it morning, afternoon or evening, whenever I get notification I just can't resist watching it. I like the way how simple these guys explain no fancy editing, animations, sponsors, AI voice, just a dude with pen and paper sharing his knowledge.
31
@joshua_tobler · 9 months ago (edited)
Imagine if we took an LLM that was specifically trained on strategic game theory, and could fine tune itself by playing a series of iterated strategic games. And then we hooked it up to military systems like, for example, our missile defense system and let it make autonomous agentic decisions.<br><br>What could possibly go wrong?
60
@jokmenen_ · 9 months ago
Great video, great explanation and a great concept. This really inspired me, thanks!
@ezgarrth4555 · 9 months ago
Look, I have experience using Claude as a coding agent, I use it every day for my job (they insist upon it, even). It is because I have that experience that I would never let an AI agent act unsupervised like you propose I should intrinsically want
174
@Locut0s · 9 months ago
I’m increasingly more worried about the unintended side effects of AI that have nothing at all really to do with constraints or alignment. There is nothing “wrong” with the attention algorithms that are behind much of the net right now before AI. All they are designed to do is keep people on a site. Unintended side effect, radicalization, social fabric fraying at the edges AI is likely, and already is, accelerating this and that’s when they are working perfectly.
16
@HowieStephens · 4 days ago
These somewhat older videos are actually super interesting to watch a year later 🫠
@AnthonySennett · 9 months ago
I used the GH agent recently and it created and included a binary in the PR. When I pointed it out it silently removed it, but I have no idea where it got that binary from. The code change was pure business logic and associated tests.
37
@b3night3d · 9 months ago
That's on you for giving it sudo access.
3
@anthonyhart7878 · 9 months ago
It can run long commands but its reading logs like crazy which fill up the context and eat up your tokens...
@MrTightFilms · 9 months ago
Teams notification at <a href="https://www.youtube.com/watch?v=JAcwtV_bFp4&t=502">8:22</a> got me
Up next

Implementing Undo - Computerphile
Computerphile · 91K views

No Regrets - What Happens to AI Beyond Generative? - Computerphile
Computerphile · 194K views

Welcome to the J-Space: Anthropic's New Technique for LLM Interpretability
Arivu · 15K views

Shor's Algorithm for Quantum Computing - Computerphile
Computerphile · 163K views

Did AI pick his pocket? - Numberphile
Numberphile · 242K views

NeRF and Gaussian Splatting - easily explained
Giulio Federico · 17K views

Anthropic Took On the Riemann Hypothesis. Here’s What Actually Happened
Ellie Sleightholm · 496K views

The Next Big SHA? SHA3 Sponge Function Explained - Computerphile
Computerphile · 167K views

Stop Button Solution? - Computerphile
Computerphile · 499K views

What the Labs Kept Secret: The German Wiki & RubyGems Hacks - Computerphile
Computerphile · 177K views

This Device Turns Anything into a Suction Cup
Steve Mould · 157K views

Why AI Tokens are so Expensive - Computerphile
Computerphile · 820K views

Philosopher David Chalmers asks: When we talk to AI, what are we talking to?
UC Berkeley and Berkeley Arts & Humanities · 205K views

AI Isn't as Powerful as We Think | Hannah Fry
New Scientist · 1.3M views

What is Bootstrapping Anyway? - Computerphile
Computerphile · 184K views

Bill Gates: A.I. ‘Makes Nuclear Weapons Look Like Nothing’ | The Ezra Klein Show
The Ezra Klein Show and 2 more · 652K views

How Watermarks Track AI Generated Content - Computerphile
Computerphile · 496K views

Ai Will Try to Cheat & Escape (aka Rob Miles was Right!) - Computerphile
Computerphile · 339K views

I'm done coding with AI
Brett Codes · 1M views

The Problem with A.I. Slop! - Computerphile
Computerphile · 715K views

The Glimmer in our Eye (with Grant Sanderson) - Numberphile Podcast
Numberphile2 · 42K views

Network Basics - Transport Layer and User Datagram Protocol Explained - Computerphile
Computerphile · 43K views

Llama.cpp vs vLLM: Which Local LLM Engine Actually Scales?
IBM Technology · 95K views

This is What Really Happens to Russia When Putin Loses | Sarah Paine
Pyotr Kurzin | Geopolitics and Pyotr Kurzin | Clips · 794K views

Generative AI's Greatest Flaw - Computerphile
Computerphile · 611K views

Yann LeCun on What Comes After LLMs
Unsupervised Learning: With Jacob Effron · 401K views