LLM Hacking

LLM Hacking refers to the exploitation of vulnerabilities in Large Language Models (LLMs) and AI-driven systems to facilitate malicious activities, ranging from data exfiltration to automated attack orchestration. This concept encompasses both direct manipulation of model outputs and the use of AI as a force multiplier for traditional cyber threats.

Key Threat Vectors

References