Ads skipped

The Hard Problem of Controlling Powerful AI Systems - Computerphile

62K views · Dec 4, 2025 · Education

Comments · 333

  • @davidswinstead · 9 months ago

    You ask an agent to find a machine's IP address, go for coffee, and when you come back the agent has SSHed in and upgraded the operating system..... my guy you are WAY TOO CHILL about this haha

    202

  • @ricoreyes6044 · 9 months ago

    It's like the Simpsons episode where Homer puts the drinking bird at the computer - "Why did I leave you in charge?!"

    81

  • @BobbyKingOfTheSeas · 9 months ago

    Be it morning, afternoon or evening, whenever I get notification I just can't resist watching it. I like the way how simple these guys explain no fancy editing, animations, sponsors, AI voice, just a dude with pen and paper sharing his knowledge.

    31

  • @joshua_tobler · 9 months ago (edited)

    Imagine if we took an LLM that was specifically trained on strategic game theory, and could fine tune itself by playing a series of iterated strategic games. And then we hooked it up to military systems like, for example, our missile defense system and let it make autonomous agentic decisions.<br><br>What could possibly go wrong?

    60

  • @jokmenen_ · 9 months ago

    Great video, great explanation and a great concept. This really inspired me, thanks!

  • @ezgarrth4555 · 9 months ago

    Look, I have experience using Claude as a coding agent, I use it every day for my job (they insist upon it, even). It is because I have that experience that I would never let an AI agent act unsupervised like you propose I should intrinsically want

    174

  • @Locut0s · 9 months ago

    I’m increasingly more worried about the unintended side effects of AI that have nothing at all really to do with constraints or alignment. There is nothing “wrong” with the attention algorithms that are behind much of the net right now before AI. All they are designed to do is keep people on a site. Unintended side effect, radicalization, social fabric fraying at the edges AI is likely, and already is, accelerating this and that’s when they are working perfectly.

    16

  • @HowieStephens · 4 days ago

    These somewhat older videos are actually super interesting to watch a year later 🫠

  • @AnthonySennett · 9 months ago

    I used the GH agent recently and it created and included a binary in the PR. When I pointed it out it silently removed it, but I have no idea where it got that binary from. The code change was pure business logic and associated tests.

    37

  • @b3night3d · 9 months ago

    That&apos;s on you for giving it sudo access.

    3

  • @anthonyhart7878 · 9 months ago

    It can run long commands but its reading logs like crazy which fill up the context and eat up your tokens...

  • @MrTightFilms · 9 months ago

    Teams notification at <a href="https://www.youtube.com/watch?v=JAcwtV_bFp4&amp;t=502">8:22</a> got me

Up next

LIVE

Implementing Undo - Computerphile

Computerphile · 91K views

LIVE

No Regrets - What Happens to AI Beyond Generative? - Computerphile

Computerphile · 194K views

LIVE

Welcome to the J-Space: Anthropic's New Technique for LLM Interpretability

Arivu · 15K views

LIVE

Shor's Algorithm for Quantum Computing - Computerphile

Computerphile · 163K views

LIVE

Did AI pick his pocket? - Numberphile

Numberphile · 242K views

LIVE

NeRF and Gaussian Splatting - easily explained

Giulio Federico · 17K views

LIVE

Anthropic Took On the Riemann Hypothesis. Here’s What Actually Happened

Ellie Sleightholm · 496K views

LIVE

The Next Big SHA? SHA3 Sponge Function Explained - Computerphile

Computerphile · 167K views

LIVE

Stop Button Solution? - Computerphile

Computerphile · 499K views

LIVE

What the Labs Kept Secret: The German Wiki & RubyGems Hacks - Computerphile

Computerphile · 177K views

LIVE

This Device Turns Anything into a Suction Cup

Steve Mould · 157K views

LIVE

Why AI Tokens are so Expensive - Computerphile

Computerphile · 820K views

LIVE

Philosopher David Chalmers asks: When we talk to AI, what are we talking to?

UC Berkeley and Berkeley Arts & Humanities · 205K views

LIVE

AI Isn't as Powerful as We Think | Hannah Fry

New Scientist · 1.3M views

LIVE

What is Bootstrapping Anyway? - Computerphile

Computerphile · 184K views

LIVE

Bill Gates: A.I. ‘Makes Nuclear Weapons Look Like Nothing’ | The Ezra Klein Show

The Ezra Klein Show and 2 more · 652K views

LIVE

How Watermarks Track AI Generated Content - Computerphile

Computerphile · 496K views

LIVE

Ai Will Try to Cheat & Escape (aka Rob Miles was Right!) - Computerphile

Computerphile · 339K views

LIVE

I'm done coding with AI

Brett Codes · 1M views

LIVE

The Problem with A.I. Slop! - Computerphile

Computerphile · 715K views

LIVE

The Glimmer in our Eye (with Grant Sanderson) - Numberphile Podcast

Numberphile2 · 42K views

LIVE

Network Basics - Transport Layer and User Datagram Protocol Explained - Computerphile

Computerphile · 43K views

LIVE

Llama.cpp vs vLLM: Which Local LLM Engine Actually Scales?

IBM Technology · 95K views

LIVE

This is What Really Happens to Russia When Putin Loses | Sarah Paine

Pyotr Kurzin | Geopolitics and Pyotr Kurzin | Clips · 794K views

LIVE

Generative AI's Greatest Flaw - Computerphile

Computerphile · 611K views

LIVE

Yann LeCun on What Comes After LLMs

Unsupervised Learning: With Jacob Effron · 401K views

YouTube, with the door locked.

Aegis plays a clean stream instead of YouTube's player, so pre-roll ads, trackers, and fingerprinting never ride along. Drop Shields any time if you want the official player back.