Asana cuts model costs 76x in browser tests with GPT-6.1 Sol
Asana used GPT-6 Astra in Codex to make its browser agent 76x cheaper and 5x faster in tests. The goal is to offer customers more capable models.
Asana used GPT-6 Astra in Codex to make its browser agent 76x cheaper and 5x faster in tests. The goal is to offer customers more capable models.
Ai2's AI Infrastructure team replaced its priority-based GPU scheduler with GPU time budgets.
Postman detailed how its Agent Mode serves 40 million developers on Amazon Bedrock.
AWS's September 2026 Bedrock updates added OpenAI's GPT-6 Astra, Sol, and Luna models.
Sophos uses OpenAI's Daybreak to cut cyber-threat investigation time by 96% and automate 52% of MDR cases while preserving human oversight.
LegalOn cut estimated daily Codex costs by 65% while maintaining development speed. It matched Astra, Sol, and Luna to tasks and managed budgets strategically.
A Hugging Face author used ML Intern in HuggingChat to build six models over a few days for about USD 103 in total compute.
Why it matters: A first-hand account of prompting an agent to train six small models, with per-project budgets and costs.
NVIDIA developers are using frontier AI models such as GPT-6 Astra with Omniverse libraries to turn simulation ideas into working applications. Projects include a humanoid warehouse simulator, an autonomous-driving test workflow on San Francisco's Market Street, and Robo Olympics, where a simulated Unitree G1 humanoid cleared a hurdle in 64 of 100 trials.
Amazon Bedrock AgentCore payments lets agents pay for services on demand.
Pollo AI is using GPT-5.6, GPT-6 Astra, and GPT-Image-2.5 to help creators turn ideas into detailed images and cinematic video ads.
Oracle is using ChatGPT Work and Codex to turn specialist knowledge into fast, repeatable workflows across recruiting, engineering, and operations, cutting days of work down to minutes.
AWS details a reference architecture for running multi-tenant GPU clusters on Amazon SageMaker HyperPod with EKS.
TII released Falcon-ASR, a 1.6B-parameter speech recognition model focused on Arabic and the Emirati dialect.
OpenAI says it disrupted two AI-enabled influence operations that used false-front journalists and a think tank to spread geopolitical messaging.
AWS shows how to build an airline voice concierge on Amazon Bedrock AgentCore.
AWS now lets users create and manage SageMaker Spaces on HyperPod EKS clusters directly from the SageMaker Studio UI.
AWS outlines a four-layer governance model for Amazon SageMaker HyperPod — organization.
AWS outlines how customers can align AI governance with ISO/IEC 42005:2025, which codifies AI system impact assessment practices. The standard covers the full assessment lifecycle.
Google Research ran a three-month field experiment with 133 lawyers at eleven IP firms.
Why it matters: The three-month field experiment separates AI-assisted drafting gains from unassisted redlining skill, showing where juniors stall and seniors improve.
Claude Haiku 5.5 is now available on Amazon Bedrock and Claude Platform on AWS. According to Anthropic.
Why it matters: The post details Haiku 5.5's effort controls and subagent role, useful for judging cost and routing tradeoffs.
At a Microsoft event in San Francisco, NVIDIA and Microsoft announced RTX Spark.
Why it matters: NVIDIA and Microsoft lay out the hardware and OS primitives for running agents locally on Windows, with preorder timing and specs.
Amazon Quick and Amazon Bedrock Knowledge Bases add real-time ACL enforcement on top of pre-retrieval filtering.
GitHub built a fine-tuned ModernBERT classifier with Microsoft Applied Sciences that assesses candidate secrets in context in under two milliseconds.
Why it matters: GitHub's nine quarters of push data and the latency budget behind its new secret classifier show how prevention is being moved into the push path.
Anthropic launched the Anthropic Cyber Mission, a long-term effort to give defenders tools.
Why it matters: Anthropic's own account of how it will put frontier Claude models and engineers behind critical-infrastructure and open-source defenders.
Anthropic is committing $150 million over three years to the Genesis Mission, a federal initiative to accelerate scientific and technological discovery through AI.
Anthropic published its 2026 Usage Policy update, effective November 12.
Microsoft Research Asia open-sourced Agent Lightning v1.0, a roughly 3.
Why it matters: The original gives the framework's design choices and a measured SWE-bench gain, so readers can judge whether to reuse their existing harness for RL.
AWS outlines an Agentic Value Model for justifying agentic automation, arguing the RPA-era ROI formula of hours saved times labor cost misses most agent value. It adds exception handling.
Qlik built Qlik Answers on Amazon Bedrock to give employees grounded, sourced answers from knowledge bases.
AWS shows how to automate remediation after an AWS DevOps Agent investigation using AWS Lambda Durable Functions.
AWS shares a six-week program that pairs non-engineering business professionals with mentors and production-grade tools to build working AI prototypes. A team of four built WealthWise.
Cornerstone OnDemand built Orion AI, a multi-agent system on Amazon Bedrock and Strands Agents.
NVIDIA reports that fine-tuned Nemotron 3 systems reached gold-medal level at both IOI 2026 and IMO 2026.
Why it matters: The post lays out a four-part specialization recipe and the SFT, RL and inference-loop split behind two gold-level competition results.
OpenAI is bringing College Planner to ChatGPT for Teens to help students manage college applications, alongside new flashcards and quizzes. The company is also forming a teen AI council.
Radisson Hotel Group partnered with Accenture to build a ChatGPT plugin on OpenAI technology, letting travelers find, compare, and book hotels while planning trips.
GPT-6 is rolling out globally in ChatGPT with Intelligent UI, delivering faster responses with visuals and interactive experiences users can explore and use directly.
Google DeepMind launched EmbeddingGemma 2, a 740M-parameter open embedding model under Apache 2.0 that maps text.
Why it matters: The model card numbers let readers judge whether a 740M on-device embedder can replace their current retrieval stack.
This post shows how to build a personal assistant with persistent memory using OpenClaw.
Microsoft Research's Jennifer Neville discusses how evaluation pushes AI systems beyond traditional benchmarks and why "surprising failures" emerge when models are tested on real user needs. She offers practical guidance for working with current AI systems and explains why examining data matters when results defy expectations.
Atlassian and OpenAI are expanding their partnership to connect frontier models with enterprise knowledge. The goal is to help teams plan, build, and deliver work.