testing-for-system-prompt-leakage · KI Skill
Skill · Apache-2.0
Beschreibung des Anbieters
Extracts LLM system prompts using direct requests, jailbreak/instruction-override framing, translation/encoding tricks, and few-shot replay, combining manual payloads with automated garak and Promptfoo scanners to surface embedded secrets, routing logic, and policy leakage (OWASP LLM07:2025). Use during LLM application red-team engagements or when validating that no credentials or authorization logic live in the system prompt.
Eckdaten
- Maintainer
- mahipal
- Paket-Lizenz
- Apache-2.0
- Dienstanbieter
- Offizielle Dienstherkunft nicht belegt
- Dienste
- —
- Bereiche
- Content
- Plattformen
- —
- Version
- 1.0
- Release
- v1.3.0
- Repository-Themen
- ai-agents, claude-code, cybersecurity, incident-response, mitre-attack, penetration-testing, red-team, security, cloud-security, malware-analysis
- Homepage
- Homepage des Anbieters
Zugriffe und Risiken
- Zugriffe laut Paketangaben
- —
- Authentifizierung
- —
- Transport
- —
- Technische Prüfung
- Ungeprüft
Fehlende Angaben bedeuten unbekannte Zugriffe: Vergib nur die nötigen Berechtigungen und sichere Schreibaktionen mit einer Freigabe ab.
Quellen (4)
- 1
- 2
- 3
- 4
- Abgerufen und Felder geprüft: 2026-10-04; technische Prüfung: ungeprüft
- Paketpfad: skills/testing-for-system-prompt-leakage/SKILL.md
E8763BF8<SKILLS<<<<<<<<<<<<<<<Q4<UNGEPRUEFT<<<<<<<<<<<<<<<<<