What shipped & dropped across AI labs, today.
Today
Mon · Sep 14 · 5 postsDeploying AI from pilot to production
This guide by Anthropic and Accenture shares seven considerations for taking enterprise AI from pilot to production, and the decisions leadership needs to make before the program advances.
Agentic coding is straining CI. Here’s how we scaled test impact analysis at Anthropic
Our CI job volume increased 25x over 6 months. We patched our test selection service three times before finding a sustainable solution.
Configure cost and quality in Copilot auto model selection
GitHub Copilot auto model selection now offers three tiers: efficiency, balance, and intelligence. Choose the tier that reflects how you want auto to weigh cost, quality, and response time for…
How healthcare organizations use Claude Tag
How Insight Health, Tennr, and Medallion are building human-agent teams with Claude Tag.
Claude for Financial Advisors
Financial advisors can now connect Claude to the custodians, portfolio platforms, CRMs, and planning tools they depend on, along with new skills tailored to the daily work of a financial advisor.
Friday
Fri · Sep 11 · 3 postsAdd VS Code Agents to Copilot usage metrics
GitHub Copilot usage metrics reports now include generally available metrics for activity in the dedicated VS Code Agents window, helping you measure adoption and engagement across enterprises and organizations. What’s…
Auto-resolution and analysis updates in Copilot code review
Copilot code review now resolves its own comments once you address them and writes smart commit messages for you when you apply its code suggestions. Behind the scenes, Copilot now…
Autonomous LLM post-training with Tunix on TPUs
Imagine going to sleep after writing a single Markdown specification and waking up to find that an A...
Thursday
Thu · Sep 10 · 4 postsGitHub Copilot weekly releases — September 7
This week, GitHub Copilot introduces Jira integration in Copilot app and adaptive model orchestration with Project HydraFusion in Copilot CLI. We also introduced new agent automation in Visual Studio Code…
What 1,000 small business owners taught us about AI
Lina Ochman, Head of U.S. SMB at Anthropic, shares what she learned during our Claude Small Business Tour, and what’s next for the program.
MAI-Code-1-Flash deprecated
We have deprecated MAI-Code-1-Flash across all GitHub Copilot experiences (including Copilot Chat, inline edits, ask and agent modes, and code completions) today, September 10, 2026. Model Deprecation date Suggested alternative…
T. Rowe Price brings more of Claude to its investment process
How T. Rowe Price is using Claude across the business, from driving fundamental research to building investment tools.
Wednesday
Wed · Sep 9 · 3 postsEnterprise managed permissions for GitHub Copilot agent operations
If you administer GitHub Copilot Business or GitHub Copilot Enterprise, you can now centrally control which agent operations are blocked, require human approval, or can proceed without a prompt. Managed…
The Anatomy of Harness Engineering: How to Evaluate, Iterate, and Guard AI Coding Agents
While end-to-end benchmarks like SWE-bench provide broad performance scores for AI agents, they are often expensive, slow, and lack the root-cause diagnostics needed to explain exactly where an agent's logic broke down. To solve this, developers should adopt behavioral evaluations—fast, local, unit-style tests that assert on discrete intermediate actions, such as verifying specific tool calls or file modifications rather than final string equality. By building these inexpensive micro-checks alongside macro benchmarks, engineering teams can confidently iterate on system prompts and upgrade models without the risk of regressions.
Announcing ADK for Kotlin 1.0: Building Production-Ready AI Agents in Kotlin, Android, and Beyond
Google has officially released version 1.0 of the Agent Development Kit (ADK) for Kotlin, achieving full feature parity with the Python and Java ADK cores to enable idiomatic, multi-agent AI development. Built on Kotlin Multiplatform (KMP), the framework leverages Kotlin Symbol Processing (KSP) for zero-reflection, type-safe function calling, alongside advanced orchestration capabilities like human-in-the-loop workflows and context compaction. Additionally, the release introduces a robust suite of Android-first extensions, allowing mobile developers to integrate local models via LiteRT-LM, cloud reasoning through Firebase AI, session persistence using Room, and semantic memory powered by AppSearch.
Tuesday
Tue · Sep 8 · 2 postsEnterprise-managed sandbox in Copilot for JetBrains
This update brings support for enterprise-managed sandbox policies, cross-file cursor jumps for next edit suggestions, global project context in chat, enterprise policy diagnostics, and a new connection between terminal Copilot…
Reducing cost and improving performance with Claude Platform
Tuning prompt caching, instructions, and effort can reduce Claude's cost without sacrificing application performance.
September 4
Fri · Sep 4 · 3 postsGitHub Copilot weekly releases — August 31
This week, GitHub Copilot expands model choice while VS Code adds new ways to manage agent sessions and get pull requests merge-ready.
GPT-6 Astra is generally available in GitHub Copilot
GPT-6 Astra from OpenAI is now available in GitHub Copilot. OpenAI’s latest general-purpose model, GPT-6 Astra, is designed for long-horizon, autonomous coding and agentic tasks. In our internal testing, GPT-6…
Setting Grok Bot loose on procurement
We gave Grok Bot access to vendor spend, contracts, and usage data. It found more than $100,000 in direct savings.