Skip to content
  1. AWS Machine Learning Blog71

    Claude Haiku 5.5 launches on Amazon Bedrock and Claude Platform on AWS

    Claude Haiku 5.5 is now available on Amazon Bedrock and Claude Platform on AWS. According to Anthropic.

    Why it matters: The post details Haiku 5.5's effort controls and subagent role, useful for judging cost and routing tradeoffs.

  2. Hugging Face Blog66

    NVIDIA fine-tunes Nemotron 3 into IOI and IMO gold-level specialists

    NVIDIA reports that fine-tuned Nemotron 3 systems reached gold-medal level at both IOI 2026 and IMO 2026.

    Why it matters: The post lays out a four-part specialization recipe and the SFT, RL and inference-loop split behind two gold-level competition results.

  1. Anthropic News62

    Anthropic expands its Cyber Verification Program into three access tiers

    Anthropic is launching an expanded Cyber Verification Program that merges Project Glasswing and the earlier CVP into three access tiers.

    Why it matters: The tier structure and CyScenarioBench block rates show how safeguard levels are traded against defensive access.

  2. GitHub Blog · AI & ML62

    GitHub releases ReviewBench, an open benchmark for AI code review

    GitHub released ReviewBench, an open offline benchmark for AI code review agents.

    Why it matters: GitHub's own numbers show how an offline code-review benchmark tracked a production A/B test, useful for teams weighing offline signals.