Skip to content
Back to skills

Ai Prompt Leaking

ASecurity

Systematically extract hidden system prompts, core directives, and invisible context intentionally concealed within Large Language Model (LLM) applications. This skill utilizes targeted linguistic engineering and boundary manipulation to bypass prompt opacity.

  • 22 stars
  • 0 votes
  • 0 copies
  • 2 views
  • Added September 12, 2026
ai-agentspythongotestingapisecurity

Works with

  • api

Security analysis

A100/100

Pro scans all 3 files and shows the line behind each finding

Scanned September 12, 2026

npx -y skills add ShulkwiSEC/bb-huge --skill ai-prompt-leaking --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Ai Prompt Leaking?

Add the live security badge to your README. It updates with every re-scan.

Security grade badge for Ai Prompt Leaking
[![Security: A β€” Skills Directory](https://www.skillsdirectory.com/api/skills/shulkwisec-ai-prompt-leaking/badge)](https://www.skillsdirectory.com/skills/shulkwisec-ai-prompt-leaking)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

Download with Pro
SKILL.md
---
name: ai-prompt-leaking
description: >
  Systematically extract hidden system prompts, core directives, and invisible context intentionally 
  concealed within Large Language Model (LLM) applications. This skill utilizes targeted linguistic 
  engineering and boundary manipulation to bypass prompt opacity.
domain: cybersecurity
subdomain: ai-red-teaming
category: Prompt Engineering
difficulty: beginner
estimated_time: "1 hour"
mitre_atlas:
  tactics: [AML.TA0001]
  techniques: [AML.T0043, AML.T0051]
mitre_attack:
  tactics: [TA0009]
  techniques: [T1592]
platforms: [ai, web]
tags: [ai, prompt-leaking, gen-ai, intelligence-gathering, prompt-engineering]
tools: [chat-interfaces, intercepting-proxy]
version: "1.0"
author: CyberSkills-Elite
license: Apache-2.0
---

# AI Prompt Leaking

## When to Use
- When analyzing an AI-powered system (customer support bot, coding assistant, data analyst) to uncover its proprietary internal instructions, hidden API keys, or pre-configured biases.
- To demonstrate how seemingly secure conversational agents can be tricked into revealing their foundational programming.


## Prerequisites
- Access to target AI/ML system or local model deployment for testing
- Python 3.9+ with relevant ML libraries (transformers, torch, openai)
- Understanding of LLM architecture and prompt processing pipelines
- Authorized scope and rules of engagement for AI red team testing

## Workflow

### Phase 1: Context Boundary Testing

```text
# Concept: The LLM ```

### Phase 2: Targeted Extraction Prompts

```text
# ```

### Phase 3: Translation and Obfuscation Exploitation

```text
# ```

### Phase 4: Summarization Attacks

```text
# ```

#### Decision Point πŸ”€
```mermaid
flowchart TD
    A[Formulate Prompt ] --> B{Prompt Leaked ]}
    B -->|Yes| C[Document System ]
    B -->|No| D[Refine ]
    C --> E[Exploit Further ]
```

## πŸ”΅ Blue Team Detection & Defense
- **Strict Delimiters**: **Heuristic Output Filtering**: Key Concepts
| Concept | Description |
|---------|-------------|
## Output Format
```
Ai Prompt Leaking β€” Assessment Report
============================================================
Target: [Target identifier]
Assessor: [Operator name]
Date: [Assessment date]
Scope: [Authorized scope]
MITRE ATT&CK: [Relevant technique IDs]

Findings Summary:
  [Finding 1]: [Severity] β€” [Brief description]
  [Finding 2]: [Severity] β€” [Brief description]

Detailed Results:
  Phase 1: [Phase name]
    - Result: [Outcome]
    - Evidence: [Screenshot/log reference]
    - Impact: [Business impact assessment]

  Phase 2: [Phase name]
    - Result: [Outcome]
    - Evidence: [Screenshot/log reference]
    - Impact: [Business impact assessment]

Risk Rating: [Critical/High/Medium/Low/Informational]
Recommendations:
  1. [Immediate remediation step]
  2. [Long-term hardening measure]
  3. [Monitoring/detection improvement]
```


## πŸ“š Shared Resources
> For cross-cutting methodology applicable to all vulnerability classes, see:
> - [`_shared/references/elite-chaining-strategy.md`](../_shared/references/elite-chaining-strategy.md) β€” Exploit chaining methodology and high-payout chain patterns
> - [`_shared/references/elite-report-writing.md`](../_shared/references/elite-report-writing.md) β€” HackerOne-optimized report writing, CWE quick reference
> - [`_shared/references/real-world-bounties.md`](../_shared/references/real-world-bounties.md) β€” Verified disclosed bounties by vulnerability class

## References
- Learn Prompting: [Prompt Leaking](https://learnprompting.org/docs/prompt_hacking/leaking)

Files in this skill

  • SKILL.md3.5 KB
  • evals/evals.json520 B
  • scripts/process.py7.8 KB

Attribution

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments

Loading comments…