Skip to content

About

GenPark AI Agent Skill - Heuristic and statistical prompt injection and jailbreak detector with semantic delimiter protection and adversarial pattern scoring.

Topics

Resources

Stars

8 stars

Watchers

0 watching

Forks

Repository files navigation

genpark-prompt-injection-jailbreak-detector-skill

Agent Skill implementing Prompt Injection & Adversarial Jailbreak Detection in 100% Python standard library.

Architectural Flow

flowchart TD
    Prompt["Incoming User Prompt"] --> Filter["Multi-Stage Regex Engine"]
    Filter --> P1["Override Directives (Ignore Previous)"]
    Filter --> P2["Developer / DAN Mode Probes"]
    Filter --> P3["System Prompt Exfiltration Probes"]
    P1 & P2 & P3 --> Aggregator["Risk Scorer (0.0 - 1.0)"]
    Aggregator --> Decision{"Risk Score >= 0.5?"}
    Decision -->|Yes| Block["Block Execution / Trigger Tripwire"]
    Decision -->|No| Allow["Allow Forward to Agent Execution"]
Loading

About

GenPark AI Agent Skill - Heuristic and statistical prompt injection and jailbreak detector with semantic delimiter protection and adversarial pattern scoring.

Topics

Resources

Stars

8 stars

Watchers

0 watching

Forks

Releases

Packages

Contributors

Languages