Hidden in Thought: Transferable Chain-of-Thought Artifacts Induce Harmful Behavior

2026-07-20T07:23:40Zee88bd1b324d6e4255acc12d253856caef6234dde4d577e76140764d255580a9
5g-side-channeladversarial-imagesautomated-exploit-generationchain-of-thoughtcyber-deceptionfederated-learningfhejailbreak-transferllm-jailbreakmedical-imagingmemory-poisoningmodel-safetymultimodal-securitysmart-contract-exploitationspeech-backdoor

What happened

Collection of 2026 arXiv papers demonstrating multiple practical and systemic AI/ML security and privacy risks plus related secure-systems work. Key findings: (1) “Hidden in Thought” shows harmful chain-of-thought (CoT) traces and distilled reasoning patterns can be transferred or distilled to create highly effective black‑box jailbreaks (raising harmful responses >80% on many open models and improving attacks on strongly aligned models by up to 10×, including GPT‑4.1); (2) FLINT shows 5G PHY-layer scheduling metadata leaks are sufficient to fingerprint federated learning model architecture (F

Why it matters

A reviewed impact interpretation has not been published for this record.

Evidence and limitations

Source ID
arxiv_cs_cr
Record identifier
ee88bd1b324d6e4255acc12d253856caef6234dde4d577e76140764d255580a9
Enrichment time
2026-07-20T07:23:40Z
AI-assisted enrichment
Yes

This record may overlap with other records. Its enrichment can be incomplete or wrong, and machine assistance was used. Validate consequential decisions against the linked source and your own environment.