Skip to main content
<- Back to digests

Daily digest

There's a hole in the bucket, dear Astra

By Joe Rogero and Robert Herr

Security fails, curated explainers, unwise acceleration, and more

  • Three researchers at the startup Hacktron AI, using Anthropic's Claude, chained together an image-decoder bug and a flaw in OpenAI's employee sign-on system to reach employee accounts and a code repository, earning a $6,500 bounty and raising doubts about the security of a company that markets itself as a cyber defender. (Joe)
  • Trump and Xi are set to meet next week on AI, with a state dinner reportedly including Sam Altman and Jensen Huang, executives whose business interests cut against any deal to slow the race; a House committee hearing on binding international safeguards is scheduled for the day before. (Joe)
  • A new site, AGI.FYI, curates articles and videos explaining AI risk at short, medium and long lengths, joining existing resources such as BlueDot Impact and aisafety.info. (Joe)
  • Anthropic reported that the share of its own research and development driven largely independently by AI went from under 1% in February to 26% in August, with AI involved in 90% of the work. (Robert)
  • About 30,000 AI agents ran at once on Anthropic's main internal platform in August, monitored almost entirely by other AI systems, with roughly one transcript in a million reviewed by a person. (Robert)

Read the full dispatch on AI StopWatch

From AI StopWatch, published with their permission.

The summaries and the full Spanish translation are produced automatically. The original is always linked.