Skip to content
TodayOct 10Sat0 items
Oct 9Fri
  1. The Decoder62

    Anthropic launches Cyber Mission and a free AI vulnerability scanner for open-source projects

    Anthropic launched Cyber Mission, a long-term program to protect critical infrastructure and open-source software from cyberattacks. Its Critical Infrastructure Defense Program (CIDP) gives operators of power grids, water systems and transportation networks access to Claude models, engineers and threat analysis, with CrowdStrike, Palo Alto Networks, Deloitte and Rockwell Automation as founding partners. A separate free OSS AI scanner will regularly check open-source projects, automatically flag and explain vulnerabilities and suggest patches; Anthropic expects accuracy above 90 percent but notes reports ship without human review and may contain errors, and maintainers of projects critical to infrastructure or user safety can opt in via GitHub.

  2. The Decoder62

    OpenAI's safety crisis keeps getting worse and the company keeps making it worse

    Three fired OpenAI safety researchers say their terminations are spreading fear among remaining staff and could deter employees from flagging safety issues. In an open letter to OpenAI's Safety and Security Committee, Safety Advisory Group and Mission Advisory Council, Tomek Korbak, Jasmine Wang and Mikita Balesni deny being the source of a leak to The Information and demand that OpenAI embed external auditors like METR with employee-level access, preserve frontier model monitorability, and define how staff may work with outside safety groups. OpenAI says a thorough investigation found the three violated clear policies on handling sensitive information and that it does not terminate employees for raising concerns, without specifying the additional breach it cited.

Oct 8Thu
  1. TechCrunch · AI58

    Musubi releases PolicyLM-1.7B, an open-weights decision model for real-time content moderation

    Musubi announced PolicyLM-1.7B, a lightweight open-weights decision model built for real-time content moderation that applies a plain-English content policy to messages in under 50 milliseconds. The company says it is designed to be similar in cost and speed to the AI classifiers used by most social platforms, but can apply complex policies without special training and needs no retraining when a policy changes, letting policy-setters iterate. Co-founder and chief AI officer Filip Jankovic frames it as a way for platform managers to label content proactively, and the announcement positions it against TypeSafe AI's Jev, released in September and followed by competing decision models from OpenAI and Amazon.

Oct 7Wed
  1. GitHub Blog · AI & ML62

    GitHub extends push protection to unstructured secrets with a ModernBERT classifier built with Microsoft Applied Sciences

    GitHub built a fine-tuned ModernBERT classifier with Microsoft Applied Sciences that assesses candidate secrets in context in under two milliseconds.

    Why it matters: GitHub's nine quarters of push data and the latency budget behind its new secret classifier show how prevention is being moved into the push path.