Asana cuts model costs 76x in browser tests with GPT-6.1 Sol
Asana used GPT-6 Astra in Codex to make its browser agent 76x cheaper and 5x faster in tests. The goal is to offer customers more capable models.
Asana used GPT-6 Astra in Codex to make its browser agent 76x cheaper and 5x faster in tests. The goal is to offer customers more capable models.
Ai2's AI Infrastructure team replaced its priority-based GPU scheduler with GPU time budgets.
Postman detailed how its Agent Mode serves 40 million developers on Amazon Bedrock.
AWS's September 2026 Bedrock updates added OpenAI's GPT-6 Astra, Sol, and Luna models.
Sophos uses OpenAI's Daybreak to cut cyber-threat investigation time by 96% and automate 52% of MDR cases while preserving human oversight.
LegalOn cut estimated daily Codex costs by 65% while maintaining development speed. It matched Astra, Sol, and Luna to tasks and managed budgets strategically.
NVIDIA developers are using frontier AI models such as GPT-6 Astra with Omniverse libraries to turn simulation ideas into working applications. Projects include a humanoid warehouse simulator, an autonomous-driving test workflow on San Francisco's Market Street, and Robo Olympics, where a simulated Unitree G1 humanoid cleared a hurdle in 64 of 100 trials.
Amazon Bedrock AgentCore payments lets agents pay for services on demand.
Oracle is using ChatGPT Work and Codex to turn specialist knowledge into fast, repeatable workflows across recruiting, engineering, and operations, cutting days of work down to minutes.
AWS details a reference architecture for running multi-tenant GPU clusters on Amazon SageMaker HyperPod with EKS.
AWS shows how to build an airline voice concierge on Amazon Bedrock AgentCore.
AWS now lets users create and manage SageMaker Spaces on HyperPod EKS clusters directly from the SageMaker Studio UI.
AWS outlines a four-layer governance model for Amazon SageMaker HyperPod — organization.
AWS outlines how customers can align AI governance with ISO/IEC 42005:2025, which codifies AI system impact assessment practices. The standard covers the full assessment lifecycle.
Claude Haiku 5.5 is now available on Amazon Bedrock and Claude Platform on AWS. According to Anthropic.
Why it matters: The post details Haiku 5.5's effort controls and subagent role, useful for judging cost and routing tradeoffs.
Amazon Quick and Amazon Bedrock Knowledge Bases add real-time ACL enforcement on top of pre-retrieval filtering.
GitHub built a fine-tuned ModernBERT classifier with Microsoft Applied Sciences that assesses candidate secrets in context in under two milliseconds.
Why it matters: GitHub's nine quarters of push data and the latency budget behind its new secret classifier show how prevention is being moved into the push path.
Microsoft Research Asia open-sourced Agent Lightning v1.0, a roughly 3.
Why it matters: The original gives the framework's design choices and a measured SWE-bench gain, so readers can judge whether to reuse their existing harness for RL.
AWS outlines an Agentic Value Model for justifying agentic automation, arguing the RPA-era ROI formula of hours saved times labor cost misses most agent value. It adds exception handling.
Qlik built Qlik Answers on Amazon Bedrock to give employees grounded, sourced answers from knowledge bases.
AWS shows how to automate remediation after an AWS DevOps Agent investigation using AWS Lambda Durable Functions.
AWS shares a six-week program that pairs non-engineering business professionals with mentors and production-grade tools to build working AI prototypes. A team of four built WealthWise.
Cornerstone OnDemand built Orion AI, a multi-agent system on Amazon Bedrock and Strands Agents.
This post shows how to build a personal assistant with persistent memory using OpenClaw.
NVIDIA's State of AI in Telecommunications report finds 89% of respondents say open source models and software are important to their AI strategy. NVIDIA announced the 30-billion-parameter Nemotron 3 Large Telco Model, fine-tuned by AdaptKey on open source telecom datasets, plus a full fine-tuning recipe via NeMo. SoftBank, AT&T and Indosat Ooredoo Hutchison are using open models for telecom-specific AI.
OpenAI and Ironclad are training and evaluating AI agents on complex contracting workflows to advance computer use for professional work.
NVIDIA Inception startups iSono Health, Whiterabbit.ai and Ataraxis AI are building AI applications for breast cancer imaging.
Cresta built Conductor, a natural-language agent builder on the Claude Agent SDK.
NVIDIA will offer DGX Spark with 64GB of unified memory from Acer, ASUS, Dell, Gigabyte.
Why it matters: The post gives the 64GB configuration's price, memory ceiling and two-unit clustering numbers, so readers can size local agent workloads against it.
ServiceNow CoreAI built AutoSynthData, a pipeline that turns a target model's failures and a stronger teacher's successes into new training tasks for enterprise agents. It generates tasks as system specification, user prompt, and verifier, then validates them in the environment and uses accepted samples for post-training, with the curriculum shifting toward remaining weaknesses. The pipeline is illustrated with EnterpriseOps Gym.
GPT-6 Astra Ultrafast is now available in the OpenAI API and to eligible ChatGPT Work and Codex users.
Why it matters: The post gives the speedup figure and the agent loop it targets, so readers can judge whether the latency change matters for their own tool-calling workflows.
NVIDIA argues AI factory ROI hinges on three factors: earning capacity, useful life.
Anthropic introduced mods, small TypeScript functions that change how Claude Code works by hooking into events such as tool calls.
Why it matters: The post details how mods hook Claude Code events and what admins can restrict, useful for judging control over an existing workflow.
Microsoft Research built a machine learning pipeline that forecasts geomagnetic storm risk for 66.
Barclays is expanding its Anthropic partnership to deploy Claude across its global operations.
Claude for Government is now generally available to federal and state agencies.
Why it matters: Details the FedRAMP High environment, spend caps and ATO-oriented audit controls agencies get before adopting Claude.
Anthropic's sales team built a buying agent on Claude Managed Agents (beta) that handles thousands of conversations daily.
Asana builds its AI agents on the Work Graph model, so agents take defined roles.
Mistral has opened a new hub in Munich housing research teams focused on Physics AI and Industrial AI.
H Company released Holo4, a new series of generalist computer-use agent models in two sizes.
Why it matters: The post gives the two model sizes, the interfaces they cover and the OSWorld 2.0 numbers, so readers can weigh a cheaper open-weight computer-use agent against closed frontier models.