https://awesomeagents.ai/leaderboards/jailbreak-red-team-leaderboard/
LLM Jailbreak and Red-Team Resistance Leaderboard | Awesome Agents
Apr 19, 2026 - Rankings of 14 frontier LLMs by adversarial robustness - how well they resist jailbreaks, prompt injection, and harmful-behavior elicitation across HarmBench,...
llm jailbreakred teamawesome agentsresistanceleaderboard
https://hashnode.com/posts/til-llm-jailbreak/6679a1d2f72caba3f40ced81
Discussion on "TIL: LLM Jailbreak" | Hashnode
Discussion on "TIL: LLM Jailbreak". Jailbreak in the context of LLM is manipulating the prompt to bypass restrictions set by the service provider. The 4 common...
llm jailbreakdiscussiontilhashnode
https://infosecured.ai/i/ai-security/llm-jailbreak-bypassing-security-cyberattacks/
LLM Jailbreak: Bypassing Security to Execute Hands-On Cyberattacks | InfoSecured.ai
Jan 15, 2025 - How LLMs like GPT-4 are being exploited through jailbreaks to bypass security safeguards and execute sophisticated cyberattacks. Learn the risks, real-world...
llm jailbreakhands onsecurityexecutecyberattacks
https://openreview.net/forum?id=WMwoSLAENS
AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks | OpenReview
Despite extensive pre-training in moral alignment to prevent generating harmful information, large language models (LLMs) remain vulnerable to jailbreak...
multi agentllmdefensejailbreakattacks