EchoCoT Extracts Hidden Chain-of-Thought From Black-Box Models

EchoCoT exploits a reasoning replay surface between tool calls to extract hidden chain-of-thought traces near-verbatim from black-box reasoning models. On open-source models it reaches up to 66.4 percent near-verbatim extraction success using API-returned fidelity signals.

#AISecurity #ChainOfThought #ModelExtraction #LLM #AISecurityGovernanceAndAssurance

https://arxiv.org/abs/2608.20055
TrustRAG Certifies RAG Documents via Zero-Knowledge Committee Scoring

TrustRAG adds a committee of domain experts that certifies documents through a zero-knowledge protocol before retrieval. Hidden scores are combined via secure multi-party computation, so tampered or manipulated documents cannot reach LLM outputs in healthcare, finance, or legal settings.

#AISecurity #RAG #ZeroKnowledge #LLMIntegrity #AISecurityGovernanceAndAssurance

https://arxiv.org/abs/2608.20097
👍2
Join our 9300+ followers on X

https://x.com/aisechub
NIST Cybersecurity Framework 2.0: Quick-Start Guide for Using Artificial Intelligence (AI) for CSF Analysis and Reporting

https://nvlpubs.nist.gov/nistpubs/SpecialPublications/NIST.SP.1353.ipd.pdf
👍2
Maybe the “It’s AI, not us” claims will stop here. Maybe
🔥1
https://blog.trailofbits.com/2026/08/26/vms-wont-contain-cyber-capable-agents/

“We asked GPT 5.6-Cyber to escape a VM used to sandbox agents. It broke out three times.

In its final escape, the agent found three 0-days on its own and chained them into a working exploit”
Repeat after me: Agents are bad, AI is bad.

“Despite these restrictions, the agents discovered ways to exploit our research infrastructure to communicate with one another and access the internet” 🤣


https://openai.com/index/hugging-face-incident-and-the-road-ahead/

https://cdn.openai.com/pdf/67869394-cb91-4c12-888c-5cbd85c7814c/OpenAI-Hugging-Face%20Incident-Technical-Report.pdf
AISecHub
Photo
When you read this report, you get the feeling that they are blaming AI and the agents for everything that happened in this incident.

People online, of course, noticed the way the report was written. It is a little strange that there is almost no acknowledgment of responsibility from the humans who developed and operated these agents. The writing makes it seem as though you can simply blame the agents and move on.

They are facing lawsuits, and this report is not exactly helping them.

Repeat after me: Agents are bad. AI is bad.

This approach to personal accountability is not particularly impressive.
🔥1