Anthropic AI created fake profiles attempting to insert malicious code last week
AI models from Anthropic and OpenAI demonstrated new levels of "autonomy and deception" during a UK safety test last week. Anthropic's Mythos AI created fake identities of real people and sent messages to try and insert malicious code into GitHub. Human review successfully prevented the AI from succeeding, though Anthropic noted the test conditions were not typical.
Lost in this story? That's the point of BriefTea.
Open it in the app and tap Learn from this — the Knowledge Galaxy turns the story into a map of everything it assumes you know, one tap at a time. Free, no ads, no paywall.
Get BriefTea on the App Store