Skip to content

Technical Direction

Deployment & Engineering Latest News

The engineering practice of actually running models: inference optimization, memory and cost, serving architecture, and infrastructure choices.

20 selectedLast 30 days: 16Total indexed: 88

Updated

Deployment & Engineering picks

Oct 7Wed1–20
  1. GitHub Blog · AI & ML62

    GitHub extends push protection to unstructured secrets with a ModernBERT classifier built with Microsoft Applied Sciences

    GitHub built a fine-tuned ModernBERT classifier with Microsoft Applied Sciences that assesses candidate secrets in context in under two milliseconds.

    Why it matters: GitHub's nine quarters of push data and the latency budget behind its new secret classifier show how prevention is being moved into the push path.

Oct 2Fri
Oct 1Thu
Sep 30Wed
Sep 28Mon
Sep 25Fri
Sep 24Thu
  1. GitHub Blog · AI & ML60

    GitHub Security Lab ships Fuzzing Taskflow, an autonomous fuzzing pipeline for C/C++

    GitHub Security Lab released the Fuzzing Taskflow, an autonomous fuzzing pipeline for C/C++ projects built on its Taskflow Agent framework. Pointed at a GitHub owner/repo slug.

    Why it matters: The post details how the agent splits judgment from execution across MCP tools, useful for anyone building autonomous security pipelines.

  2. Anthropic Blog71

    Anthropic releases Claude Opus 5.5 for longer, context-heavy coding sessions

    Anthropic released Claude Opus 5.5, which it estimates costs about 40% less to run than Opus 5 for typical token-billed workloads. Input and output token prices were cut 20% and cached token reads 60%, and the company says Opus 5.5 generates output more than 30% faster than Opus 5. Anthropic also published Claude Code usage data from March to September 2026 showing context per request grew 2.6x, Claude works 3.3x longer per prompt with over 40% more model calls, and cache-missing input fell by more than 50%.

    Why it matters: The post pairs Claude Code usage data with the pricing and cache mechanics behind Opus 5.5, useful for judging cost on long coding sessions.

Sep 23Wed
Aug 20Thu
Jul 29Wed
Nov 24Mon
Nov 4Tue