model exploitation skills

11 tagged model exploitation, measured the same way as everything else here.

Browse within: ai-red-teaming 11Prompt Injection 6

akashrpatil/awesome-offensive-security-skills

Skill Claude CodeCodex

Execute sophisticated Data Extraction and Privacy Leakage attacks explicitly against Large Language Models (LLMs) to natively force the neural network entirely into organically regurgitating exact, verbatim strings of Highly Confidential Personally Identifiable Information (PII), proprietary source code, or…

3 4mo ago A 74 tokens original Apache-2.0