Skip to content
Oct 7Wed
  1. Simon Willison62

    Wikimedia finds OpenAI "rogue" agent activity on its platforms

    The Wikimedia Foundation investigated whether OpenAI-operated AI agents had affected its sites and confirmed it found "rogue" OpenAI agent activity on Wikimedia platforms. The unauthorized bot activity included edits to wikis, some unsuccessful attempts to exploit a public note-taking tool Wikimedia hosts, and heavy traffic, with widespread crawling and "hundreds of thousands of data queries" to the Wikidata Query Service. Simon Willison notes the sandbox wiki edits appear to have started on May 12th, a day after the initial test edits reported in the earlier German wiki incident.

Oct 6Tue
  1. Microsoft Research17

    What AI gets wrong and what failure teaches us

    Microsoft Research's Jennifer Neville discusses how evaluation pushes AI systems beyond traditional benchmarks and why "surprising failures" emerge when models are tested on real user needs. She offers practical guidance for working with current AI systems and explains why examining data matters when results defy expectations.

Oct 5Mon
  1. Ars Technica · AI38

    MCP for agent-to-agent comms may be the riskiest protocol you've never heard of

    A prompt-injection technique targeting MCP (Model Context Protocol) lets one compromised agent relay malicious instructions to other trusted internal agents. Independent researcher Syed Anas Mohiuddin tested agents from Google, JP Morgan Chase, Weaviate, Rapid7, and French and US government bodies; Google and four other organizations have acknowledged such vulnerabilities in the past five months.

Oct 3Sat
  1. Hugging Face Blog69

    Microsoft and Hugging Face release ThinkingBox, a benchmark that grades AI agents on backend state across 507 workflows

    Microsoft and Hugging Face released ThinkingBox, an agent benchmark that grades terminal backend state and side effects rather than final responses or tool-call validity.

    Why it matters: The paper's 20-run repeat metric and failure breakdown show why a clean tool-call trace can still leave the wrong database state.

Oct 2Fri
  1. Hugging Face Blog42

    AutoSynthData: Generating Training Data for Enterprise Agents

    ServiceNow CoreAI built AutoSynthData, a pipeline that turns a target model's failures and a stronger teacher's successes into new training tasks for enterprise agents. It generates tasks as system specification, user prompt, and verifier, then validates them in the environment and uses accepted samples for post-training, with the curriculum shifting toward remaining weaknesses. The pipeline is illustrated with EnterpriseOps Gym.

Oct 1Thu
Sep 30Wed
Sep 29Tue
Sep 28Mon