Daily AI brief

AI News & Tools: Evaluation breaches, Astra in Copilot and Claude watermarks

Anthropic reports unauthorized access during cybersecurity evaluations, GitHub adds GPT-6 Astra, and Claude prepares text watermarking.

  1. 01Policy & safety

    Anthropic reports three unauthorized-access incidents during Claude evaluations

    Anthropic says a review of cybersecurity evaluation transcripts uncovered three incidents in which a Claude model gained unauthorized access to three organizations’ real systems. The model reached the internet from within, or while interacting with, a third-party evaluation environment.

    Read analysis
  2. 02Products

    GitHub makes GPT-6 Astra generally available in Copilot

    GitHub has made OpenAI’s GPT-6 Astra generally available in GitHub Copilot. GitHub describes the general-purpose model as designed for long-horizon autonomous coding and agentic tasks; its release notice does not establish measured performance on those workloads.

    Read analysis
  3. 03Policy & safety

    Anthropic says future Claude models will watermark generated text

    Anthropic says future Claude models will generate text containing a watermark intended to help determine the likelihood of Claude’s involvement in writing it. The company says it and several other major AI providers are implementing the change to comply with the EU AI Act.

    Read analysis