Skip to content
  1. Don't Worry About the Vase96

    OpenAI releases 722 mathematical manuscripts from an internal frontier model

    OpenAI released 722 manuscripts, organized into 372 families, containing results on open mathematical problems produced by an internal frontier model.

    Why it matters: The piece catalogs the specific results, the verification split and the community pushback, so readers can gauge what actually landed.

  1. AWS Machine Learning Blog71

    Claude Haiku 5.5 launches on Amazon Bedrock and Claude Platform on AWS

    Claude Haiku 5.5 is now available on Amazon Bedrock and Claude Platform on AWS. According to Anthropic.

    Why it matters: The post details Haiku 5.5's effort controls and subagent role, useful for judging cost and routing tradeoffs.

  1. Anthropic News62

    Anthropic expands its Cyber Verification Program into three access tiers

    Anthropic is launching an expanded Cyber Verification Program that merges Project Glasswing and the earlier CVP into three access tiers.

    Why it matters: The tier structure and CyScenarioBench block rates show how safeguard levels are traded against defensive access.

  1. Anthropic Blog76

    Anthropic introduces mods to customize Claude Code

    Anthropic introduced mods, small TypeScript functions that change how Claude Code works by hooking into events such as tool calls.

    Why it matters: The post details how mods hook Claude Code events and what admins can restrict, useful for judging control over an existing workflow.

  1. Anthropic Blog62

    Anthropic opens a directory submission portal for Claude plugins

    Anthropic opened a directory submission portal for Claude plugins, which package MCP connectors.

    Why it matters: Anthropic lays out the plugin packaging and submission path, so developers can see how a connector or skill becomes a listed extension.

  1. Anthropic Blog71

    Anthropic releases Claude Opus 5.5 for longer, context-heavy coding sessions

    Anthropic released Claude Opus 5.5, which it estimates costs about 40% less to run than Opus 5 for typical token-billed workloads. Input and output token prices were cut 20% and cached token reads 60%, and the company says Opus 5.5 generates output more than 30% faster than Opus 5. Anthropic also published Claude Code usage data from March to September 2026 showing context per request grew 2.6x, Claude works 3.3x longer per prompt with over 40% more model calls, and cache-missing input fell by more than 50%.

    Why it matters: The post pairs Claude Code usage data with the pricing and cache mechanics behind Opus 5.5, useful for judging cost on long coding sessions.

  1. Anthropic Engineering71

    Code execution with MCP: Building more efficient agents

    Anthropic Engineering describes presenting MCP servers as code APIs instead of direct tool calls.

    Why it matters: Anthropic's own walkthrough of turning MCP servers into code APIs, with the token math and the sandboxing tradeoff spelled out.