Beyond the Leaderboard: A Synthesis of Tool-Use, Planning, and Reasoning Failures in Large Language Model Agents
A synthesis of known failure modes in LLM-based agents, covering tool-use errors, planning breakdowns, and reasoning vulnerabilities that compound into systemic security risks.