Foundation-Model-Paper-Notes.
79BAP-Jailbreak-Vision-Language-Models-via-Bi-Modal-Adversarial-Prompt. Jupyter Notebook
61Jailbreak_GPT4o. Python
28RACE. Python
27ClawGuard. ClawGuard is a comprehensive security toolkit designed to mitigate risks associated with autonomous agents, such as OpenClaw and other LLM Agents.
25SafeBench. Python
22Awesome-Trustworthy-GenAI.
8DeepSeek-Safety-Eval. Python
7Open-Source-Backdoor-Learning. official/unofficial open source code/dataset for backdoor attack and defense
4ATLAS_Challenge_2025. Python
3Awesome-Datasets-for-Cybersecurity.
3Best_Practice_for_Security_of_AI. Jupyter Notebook
3VEIL. Code and data for paper VEIL: Jailbreaking Text-to-Video Models via Visual Exploitation from Implicit Language
3AegisLink. 灵盾 — AI Agent 身份与权限中枢
2PRISM-Programmatic-Reasoning-with-Image-Sequence-Manipulation-for-LVLM-Jailbreaking. official code
1personal-backup. backup
1Eval_GPT4o_on_SelectiveTest. Eval Performance of GPT-4o on Selective High School Placement Test
1K410ng-s-NG-WAF. The waf is designed by dr0ps,and remould by Yale
1Best_Practice_in_AI_for_Security. Best_Practice_in_AI_for_Security
1Awesome-material-for-TrustworthyAI.
1Evolving-Deception. Evolving Deception: When Agents Evolve, Deception Wins
1