Briefing
The fast feed: breaking AI news from the global tech press, deduplicated and time-ordered.
The fast feed: breaking AI news from the global tech press, deduplicated and time-ordered.
Latest in Briefing
30 stories
Introducing Claude Sonnet 5 on AWS: Anthropic’s most capable Sonnet model
Today, we’re excited to announce the availability of Anthropic’s most advanced Sonnet model, Claude Sonnet 5, on Amazon Bedrock and Claude Platform on AWS. Clau…

Per-user AI credit budgets available for cost centers
Enterprise admins can now set a cost center user-level budget: one per-user AI credit budget on a cost center that applies to every individual in it. As members…

Google launches Nano Banana 2 Lite for fast AI images and Gemini Omni Flash for video via API
Google adds two new generative AI models. Nano Banana 2 Lite generates images in four seconds at $0.034 a pop. Gemini Omni Flash brings video generation and edi…

OpenAI reportedly cut response costs for guest ChatGPT users by more than half
According to a report by The Information, OpenAI has cut inference costs for its AI models by more than half. The company applied the optimizations to ChatGPT, …

Anthropic launches Claude Science, an AI workspace built specifically for researchers
Anthropic released Claude Science, an AI workbench for researchers. More than 60 preconfigured skills cover fields like genomics and computational chemistry, an…
From linear gates to learning loops: Rewiring biopharma R&D with AI
An AI-powered R&D model built on connected decision loops—replacing the current sequential, fragmented workflow—can reduce uncertainty, compress timelines, and …

Claude Sonnet 5 is generally available for GitHub Copilot
Claude Sonnet 5 is Anthropic’s latest Sonnet-class model, now available in GitHub Copilot. It brings strong coding performance to everyday development and…
Claude Code v2.1.197

Implementing resilience patterns with Amazon Bedrock and LLM gateway
In this post, you will learn five practical patterns for building resilient generative AI applications on AWS, progressing from native Amazon Bedrock features t…

Simplify multi-account access to Amazon Bedrock models with managed entitlements
In this post, we show you how to use managed entitlements for Amazon Bedrock to subscribe once from a central account and distribute model access across your or…

Build generative UI for AI agents on Amazon Bedrock AgentCore with the AG-UI protocol
This post walks through how AG-UI integrates into the Fullstack AgentCore Solution Template (FAST) to build interactive agent frontends on Amazon Bedrock AgentC…

SkillOpt: Agent skills as trainable parameters
AI agents often fail because their instructions, or skills, are manually modified with no guarantee of improvement. Learn how SkillOpt turns skill editing into …
Introducing GeneBench-Pro
Introducing GeneBench-Pro, a new benchmark testing AI performance in genomics, biology, and scientific research using complex, real-world datasets.…

NVIDIA BioNeMo Agent Toolkit Brings Accelerated AI to Life Sciences Researchers in Claude Science
Life sciences has entered an era of computational scale, and for more than a decade, NVIDIA has built the full GPU-accelerated computing stack — spanning hardwa…

Build agents even faster with Gemini Enterprise Agent Platform’s fully-managed, remote MCP server
A couple of months ago, we announced that over 50 Google-managed MCP servers are available. Today, we’ll dive into how to use the Gemini Enterprise Agent Platfo…

How Schrödinger sped up molecular discovery by 4x with Alphaevolve
Computational chemistry researchers have traditionally faced a frustrating trade-off when simulating molecular interactions: use fast classical force fields tha…

Bringing speed and strong cost performance to the market with Gemini Omni Flash and Nano Banana 2 Lite
Great creative happens when your tools move at the speed of your ideas. To help you create rich, reliable experiences while reducing regeneration time and costs…

Fine-tune Amazon Nova models for accurate email data extraction
In this post, you'll learn how fine-tuning Amazon Nova models using Amazon SageMaker AI addresses these specific issues by teaching the models to recognize your…

Building bilingual NER for cargo logistics with Amazon Bedrock
In this post, we share the technical approach using token-based distillation, lessons learned, and deployment architecture. If you face similar bilingual NER ch…

How Outpost VFX Uses AWS to Accelerate AI Model Training for Visual Effects
In this post, we explore how Outpost VFX achieved 8x faster training speeds using AWS infrastructure to transform their face replacement workflow, the technical…
Copilot Agent is now available in JetBrains AI Assistant
Today, JetBrains and GitHub are announcing a deeper integration between JetBrains AI Assistant and GitHub Copilot. Millions of developers already rely on the Gi…
Core dump epidemiology: fixing an 18-year-old bug
OpenAI engineers used large-scale core dump analysis to debug rare infrastructure crashes, uncovering both a hardware fault and a long-standing software bug.…
How ChatGPT adoption has expanded
New OpenAI Signals data shows how ChatGPT adoption is growing globally, with users increasing usage, exploring more capabilities, and driving growth across regi…

San Francisco's AI boom is pricing out six-figure tech workers who can't find rent under $5,000
San Francisco's AI boom is driving up the cost of living so fast that even couples earning $365,000 a year can't find an affordable apartment. Median rent sits …

Meituan's LongCat-2.0 shows China can train massive AI models without Nvidia
Meituan trains a 1.6 trillion parameter AI model entirely on Chinese chips, no Nvidia required. The article Meituan's LongCat-2.0 shows China can train mas…

Trump Administration NEPA Reforms: A Win for All Americans
“In this Administration, NEPA’s regulatory reign of terror has ended,” said White House Council on Environmental Quality Chairman Katherine Scarlett. For decade…

How Jaiveer Singh Is Helping Robots — and Developers — Move Faster
When Jaiveer Singh talks about robots, he doesn’t begin with spectacle. He begins with infrastructure: the boards inside machines, the software that lets develo…

How NVIDIA’s Inference Software Stack Powers the Lowest Token Cost
As organizations move from AI pilots to production AI factories, infrastructure decisions have shifted from peak chip specifications to cost per token: how many…

GPT-5.6 is here but...
plus the insanity of inference…

Into the Omniverse: Three Workflows for Improving Vision AI Agent Accuracy With Synthetic Data and Fine-Tuning
Editor’s note: This post is part of Into the Omniverse, a series focused on how developers, 3D practitioners, and enterprises can transform their workflows usin…