AI Exploiting Vulnerabilities

LLM agents powered by GPT-4 can exploit unpatched vulnerabilities successfully, even those discovered after GPT-4's knowledge cutoff date. Real-world websites, software, and python packages with vulnerabilities are targeted. The agents have an 87% success rate in exploiting these vulnerabilities, showing their effectiveness in cybersecurity.