Robust AI Security and Alignment: A Sisyphean Endeavor?
12 December 2025 at 13:00
arXiv:2512.10100v1 Announce Type: new
Abstract: This manuscript establishes information-theoretic limitations for robustness of AI security and alignment by extending G\"odel's incompleteness theorem to AI. Knowing these limitations and preparing for the challenges they bring is critically important for the responsible adoption of the AI technology. Practical approaches to dealing with these challenges are provided as well. Broader implications for cognitive reasoning limitations of AI systems are also proven.