A research team says it found an API-level technique that made lower-tier models reveal hidden reasoning generated by stronger models from the same provider. The attack worked across Anthropic, OpenAI and Google in early-July tests and exposed credentials and personal data embedded in publicly shared agent logs.
Continue reading
The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.
Already purchased? Sign in✓ Signed in — this article isn’t included in your current plan.