Перейти к содержимому

Why Human Control Isn't Enough in Military AI with Heidy Khlaaf

Responsible Bytes Podcast

0:00 / 0:00

Why Human Control Isn't Enough in Military AI with Heidy Khlaaf

426 просмотров · 2 недели назад
Responsible Bytes Podcast
101 подписчик
426 просмотров · 2 недели назад
In this episode of Responsible Bytes, Dr. Zena Assaad speaks with Heidy Khlaaf, Chief AI Scientist at the AI Now Institute, about why military AI systems are far less reliable than their marketing suggests, especially in safety-critical and combat settings. They unpack the gap between “accuracy” as a narrow model metric and real-world performance, explaining how opaque, brittle systems, automation bias, outdated data, and poor interoperability can lead to serious mistakes while obscuring accountability. The conversation also explores the power dynamics behind the push for AI in warfare, including what Heidy characterises as a “safety theatre,” version-control concerns, and how companies and governments normalise speed over caution. Overall, the episode is a sharp critique of the idea that AI can safely replace human judgment in high-stakes environments, and a call for more skepticism, technical rigor, and accountability in how these systems are deployed. Chapters: 00:00 Introduction to Responsible Bytes and Heidy Khlaaf’s background 01:19 Why meaningful human control is a weak answer to unsafe AI 04:56 Human oversight cannot fix fundamentally inaccurate systems 06:14 What “accuracy” means in defense versus machine learning 07:39 Real-world accuracy collapse in military deployment 10:39 How interfaces create the illusion of intelligence 11:27 “Magic” language, data scale, and the illusion of AGI 13:37 Human feedback labor behind model fine-tuning 14:28 AI cannot know what it does not know 16:19 Why fault tolerance is hard with probabilistic systems 18:05 The Minab school bombing and bad data pipelines 19:30 Sensors, corruption, and missing interoperability 22:22 Why “just feed it more data” is not a real solution 24:13 Speed as the real military incentive behind AI 26:04 How low accuracy can obscure accountability for civilian harm 27:27 “Carpet bombing” by another name through AI-enabled targeting 30:14 Why “AI solutions” often do not solve a real capability gap 33:09 Why large language models are worse than purpose-built models 36:50 Anthropic, the US Department of Defense, and the public red lines dispute 37:47 Why the dispute is better understood as safety theater 41:42 The lack of iteration and version control in model deployment 43:15 How model providers can change behavior in real time 45:57 Concentration of power and unclear model alignment 47:40 The power shift revealed by the Anthropic and DoD standoff 49:02 Why private companies should not be monitoring state operations 51:17 The self-serving feedback loop between vendors and defense buyers 53:12 Companies grading their own homework on validation and evaluation 54:39 Heidy’s responsible byte: don’t take AI company claims at face value More about our guest: 📸Instagram:   / hak90   📘LinkedIn:   / heidy-khlaaf   🔵BlueSky: https://www.heidyk.com/ 🌐Website: https://www.heidyk.com/ More about our host: 📸Instagram:   / zena_assaad   📘LinkedIn:   / dr-zena-assaad   🔵BlueSky: https://bsky.app/profile/zenaassaad.b... 🌐Website: https://www.zenaassaad.com/ Watch or listen to more episodes: 🍎 https://podcasts.apple.com/au/podcast... 🔊 https://open.spotify.com/show/2x2yNvU... Or find us anywhere else you get your podcasts! Follow us online: 📸Instagram:   / responsiblebytespodcast   📘LinkedIn:   / responsible-bytes-podcast   🌐Website: https://www.zenaassaad.com/responsibl...