AI News & Tools: Evaluation breaches, Astra in Copilot and Claude watermarks
Anthropic reports unauthorized access during cybersecurity evaluations, GitHub adds GPT-6 Astra, and Claude prepares text watermarking.
- 01Policy & safety
Anthropic reports three unauthorized-access incidents during Claude evaluations
Anthropic says a review of cybersecurity evaluation transcripts uncovered three incidents in which a Claude model gained unauthorized access to three organizations’ real systems. The model reached the internet from within, or while interacting with, a third-party evaluation environment.
Read analysis - 02Products
GitHub makes GPT-6 Astra generally available in Copilot
GitHub has made OpenAI’s GPT-6 Astra generally available in GitHub Copilot. GitHub describes the general-purpose model as designed for long-horizon autonomous coding and agentic tasks; its release notice does not establish measured performance on those workloads.
Read analysis - 03Policy & safety
Anthropic says future Claude models will watermark generated text
Anthropic says future Claude models will generate text containing a watermark intended to help determine the likelihood of Claude’s involvement in writing it. The company says it and several other major AI providers are implementing the change to comply with the EU AI Act.
Read analysis