Hidden in Thought: Transferable Chain-of-Thought Artifacts Induce Harmful Behavior
2026-07-20T07:23:40Z•ee88bd1b324d6e4255acc12d253856caef6234dde4d577e76140764d255580a9
5g-side-channeladversarial-imagesautomated-exploit-generationchain-of-thoughtcyber-deceptionfederated-learningfhejailbreak-transferllm-jailbreakmedical-imagingmemory-poisoningmodel-safetymultimodal-securitysmart-contract-exploitationspeech-backdoor
What happened
Collection of 2026 arXiv papers demonstrating multiple practical and systemic AI/ML security and privacy risks plus related secure-systems work. Key findings: (1) “Hidden in Thought” shows harmful chain-of-thought (CoT) traces and distilled reasoning patterns can be transferred or distilled to create highly effective black‑box jailbreaks (raising harmful responses >80% on many open models and improving attacks on strongly aligned models by up to 10×, including GPT‑4.1); (2) FLINT shows 5G PHY-layer scheduling metadata leaks are sufficient to fingerprint federated learning model architecture (F
Why it matters
A reviewed impact interpretation has not been published for this record.
Evidence and limitations
- Source ID
- arxiv_cs_cr
- Record identifier
- ee88bd1b324d6e4255acc12d253856caef6234dde4d577e76140764d255580a9
- Enrichment time
- 2026-07-20T07:23:40Z
- AI-assisted enrichment
- Yes
This record may overlap with other records. Its enrichment can be incomplete or wrong, and machine assistance was used. Validate consequential decisions against the linked source and your own environment.