Ai2 replaces priority scheduler with GPU time budgets and fair-share allocation
Ai2's AI Infrastructure team replaced its priority-based GPU scheduler with GPU time budgets.
Ai2's AI Infrastructure team replaced its priority-based GPU scheduler with GPU time budgets.
Postman detailed how its Agent Mode serves 40 million developers on Amazon Bedrock.
Simon Willison shipped a Newsletters index page for his blog.
A Hugging Face author used ML Intern in HuggingChat to build six models over a few days for about USD 103 in total compute.
Why it matters: A first-hand account of prompting an agent to train six small models, with per-project budgets and costs.
NVIDIA developers are using frontier AI models such as GPT-6 Astra with Omniverse libraries to turn simulation ideas into working applications. Projects include a humanoid warehouse simulator, an autonomous-driving test workflow on San Francisco's Market Street, and Robo Olympics, where a simulated Unitree G1 humanoid cleared a hurdle in 64 of 100 trials.
AWS details a reference architecture for running multi-tenant GPU clusters on Amazon SageMaker HyperPod with EKS.
Simon Willison tested whether Claude Opus 5.5 could compose computer game music by prompting it to design a text-based music format and build a playable artifact with example tracks. The model leaned heavily into the Monkey Island theme, but Willison called the results surprisingly good. He wonders whether competent music composition is a newly emerged capability for text models, similar to recent 3D graphics advances.
Simon Willison documented running Parseable, a new OpenTelemetry-compatible observability tool.
AWS shows how to build an airline voice concierge on Amazon Bedrock AgentCore.
AWS outlines a four-layer governance model for Amazon SageMaker HyperPod — organization.
AWS outlines how customers can align AI governance with ISO/IEC 42005:2025, which codifies AI system impact assessment practices. The standard covers the full assessment lifecycle.
Amazon Quick and Amazon Bedrock Knowledge Bases add real-time ACL enforcement on top of pre-retrieval filtering.
AWS outlines an Agentic Value Model for justifying agentic automation, arguing the RPA-era ROI formula of hours saved times labor cost misses most agent value. It adds exception handling.
Qlik built Qlik Answers on Amazon Bedrock to give employees grounded, sourced answers from knowledge bases.
AWS shows how to automate remediation after an AWS DevOps Agent investigation using AWS Lambda Durable Functions.
AWS shares a six-week program that pairs non-engineering business professionals with mentors and production-grade tools to build working AI prototypes. A team of four built WealthWise.
Cornerstone OnDemand built Orion AI, a multi-agent system on Amazon Bedrock and Strands Agents.
This post shows how to build a personal assistant with persistent memory using OpenClaw.
Cresta built Conductor, a natural-language agent builder on the Claude Agent SDK.
GitHub outlines three skills developers need as AI reshapes their work: directing AI agents rather than just using them.
Anthropic's sales team built a buying agent on Claude Managed Agents (beta) that handles thousands of conversations daily.
Asana builds its AI agents on the Work Graph model, so agents take defined roles.
GitHub Copilot app's canvases are customizable, bidirectional interfaces you create by running the /create-canvas skill and describing the workflow in plain English. The agent builds the UI in the right-side panel, and both you and the agent can update its shared state at the same time. Canvases are saved as extensions for reuse or team sharing, and ready-made ones are available via Awesome Copilot.
GitHub rebuilt the pull request view in its Copilot app to keep reviews fast on enormous diffs.
Interconnects has published a reading list on open models, covering foundation topics.
Mistral helped a European energy operator migrate 40,000 lines of Fortran 77 to C++.
Anthropic Engineering describes presenting MCP servers as code APIs instead of direct tool calls.
Why it matters: Anthropic's own walkthrough of turning MCP servers into code APIs, with the token math and the sandboxing tradeoff spelled out.