MultiTurnPSB: Evaluating Multi-Turn Jailbreak Attacks an dClassifier-Based Defenses for Medical AI Safety
arXiv 2606.02630•7dc69dcaf613e4f7950e3cfcd1affdc8073a8755c98cf453b76383c091ba0acd
CREEPD-JudgeISPMLITL (Lies-in-the-Loop)LLM safetyOWASP-LLM-Top-10RA-ICAagent refusalbyte-native LLMconsent integritydeepfake detectionfalse positivesidentity securityinference-cost attackinput-side classifierjailbreak attacksjudge model manipulationknowledge poisoningmalware analysismulti-turn jailbreakoutput rewritingparaphrase brittlenessretrieval-augmented generationrobustness
Paper metadata
- arXiv ID
- 2606.02630
- Version
- Not specified by this published record
- Category
- Computer Science — Cryptography and Security (cs.CR)
The PDF link points to arxiv.org. Baitaphish does not expose a private stored PDF.
Evidence and limitations
- Source ID
- arxiv_cs_cr
- Record identifier
- 7dc69dcaf613e4f7950e3cfcd1affdc8073a8755c98cf453b76383c091ba0acd
- Enrichment time
- 2026-06-03T07:23:31Z
- AI-assisted enrichment
- Yes
This record may overlap with other records. Its enrichment can be incomplete or wrong, and machine assistance was used. Validate consequential decisions against the linked source and your own environment.