AI Sec

AI Security Topics

Every offensive AI security topic covered here: prompt injection, jailbreaks, agent and tool-use exploitation, adversarial ML, and red team methodology.

Tags

  • #llm-security 28
  • #prompt-injection 26
  • #red-team 26
  • #agent-security 14
  • #jailbreak 13
  • #adversarial-ml 12
  • #indirect-injection 4
  • #tooling 4
  • #attack-vectors 3
  • #owasp 3
  • #spoke 3
  • #chatgpt 2
  • #long-context 2
  • #membership-inference 2
  • #model-extraction 2
  • #prompt-engineering 2
  • #rag 2
  • #tool-use 2
  • #adversarial-training 1
  • #agents 1
  • #ai-red-team 1
  • #ai-security 1
  • #alignment 1
  • #application-security 1
  • #attack-techniques 1
  • #automated-attacks 1
  • #behavioral-evaluation 1
  • #bypass-techniques 1
  • #ceh 1
  • #custom-gpts 1
  • #cve 1
  • #defense 1
  • #detection 1
  • #editorial-policy 1
  • #evasion 1
  • #faq 1
  • #garak 1
  • #gcg 1
  • #governance 1
  • #gpt-4 1
  • #gpt-security 1
  • #guardrails 1
  • #hub 1
  • #insecure-output-handling 1
  • #interpretability 1
  • #knowledge-corruption 1
  • #llm-bypass 1
  • #llm-monitoring 1
  • #llm-security-vulnerabilities 1
  • #many-shot-jailbreaking 1
  • #methodology 1
  • #model-inversion 1
  • #model-theft 1
  • #multi-turn 1
  • #multimodal 1
  • #openai 1
  • #oscp 1
  • #owasp-llm01 1
  • #payload-construction 1
  • #payload-delivery 1
  • #pillar 1
  • #poisoning 1
  • #pyrit 1
  • #reporting 1
  • #scoping 1
  • #system-prompt-leakage 1
  • #taxonomy 1
  • #threat-modeling 1
  • #training-data-privacy 1
  • #transferability 1

Categories

red-team 12 posts

prompt-injection 8 posts

jailbreak 6 posts

primer 3 posts

Spoke 3 posts

Editorial 1 post

hub 1 post

Pillar 1 post