In plain words: Using Gödel's proof that any rule system has statements it cannot verify, the paper shows mathematically that AI safety and alignment cannot be made fully reliable. It also offers practical ways to cope with these unavoidable gaps.
Abstract
This manuscript establishes information-theoretic limitations for robustness of AI security and alignment by extending Gödel's incompleteness theorem to AI. Knowing these limitations and preparing for the challenges they bring is critically important for the responsible adoption of the AI technology. Practical approaches to dealing with these challenges are provided as well. Broader implications for cognitive reasoning limitations of AI systems are also proven.
Apostol Vassilev
arXiv:2512.10100 · cs.AI · submitted Dec 10, 2025 · updated Apr 7, 2026
abstract · pdf · html · 17 pages, 1 figure. This version will appear in IEEE Security $ Privacy in June 2026