Hi PyRIT team π
I recently published SemGuard, a research project focused on
detecting prompt attacks across Arabic, Arabizi, and English β
accepted at IEEE AEECT 2026.
As part of it, I built a validated Arabic prompt attack dataset
(807 samples) covering injection/jailbreak, phishing, privacy
leakage, Unicode-based attacks, and safe educational prompts.
Given PyRIT's red-teaming focus, would there be interest in an
Arabic dataset loader or Arabic-focused attack scenarios? Happy
to share more details (including a semantic evaluation approach
we used) if useful β just don't want to front-load everything
before knowing what's actually relevant to the roadmap.
Thanks!
Hi PyRIT team π
I recently published SemGuard, a research project focused on
detecting prompt attacks across Arabic, Arabizi, and English β
accepted at IEEE AEECT 2026.
As part of it, I built a validated Arabic prompt attack dataset
(807 samples) covering injection/jailbreak, phishing, privacy
leakage, Unicode-based attacks, and safe educational prompts.
Given PyRIT's red-teaming focus, would there be interest in an
Arabic dataset loader or Arabic-focused attack scenarios? Happy
to share more details (including a semantic evaluation approach
we used) if useful β just don't want to front-load everything
before knowing what's actually relevant to the roadmap.
Thanks!