Research
Curated daily papers, lab research blogs and national AI research programs across the US, China, EU, UK, Japan and beyond.
Curated daily papers, lab research blogs and national AI research programs across the US, China, EU, UK, Japan and beyond.
Latest in Research
30 stories
Exploring self-distilled reasoning for supervised fine-tuning with Amazon Nova
In this post, we explore an idea for generating thinking tokens for datasets that lack reasoning traces in SFT customization. We first examine the reasoning sup…

How Couchbase built a multi-model AI architecture for Capella iQ with Amazon Bedrock
This post describes how Couchbase adopted Amazon Bedrock to power Capella iQ with Anthropic’s Claude family of models, the architectural decisions behind their …
Safety and alignment in an era of long-horizon models
OpenAI shares lessons from deploying long-running AI models, highlighting new safety risks, observed failures, and improved safeguards through iterative deploym…

13 hands-on demos to build on Gemini Enterprise Agent Platform
Earlier this year, we introduced Gemini Enterprise Agent Platform, where you can build, scale, govern, and optimize agents. Today, we’re sharing 13 demos that w…

How Smartsheet built a remote MCP server on AWS
In this post, we cover a high-level view of the Smartsheet remote MCP architecture, with a focus on the AWS infrastructure behind it. This includes security, go…

NVIDIA Vera Rubin Maximizes Intelligence per Dollar for Post-Training Workloads – a Key Metric for Agentic AI
Lowest cost per token from extreme codesign maximizes intelligence per dollar for post-training in the agentic era.…

Google is a Leader and positioned furthest in Vision and highest in Execution in the 2026 Gartner® Magic Quadrant™ for Conversational AI Platforms
For the second consecutive year, Google has been named a Leader in the Gartner® Magic Quadrant™ for Conversational AI Platforms. Google received the furthest an…
What the 2026 World Cup reveals about the next generation of sports fans
This year’s tournament offers a window into consumers’ changing behaviors, desire for community, and focus on identity—especially among Latino fans. Are busines…

NVIDIA Introduces New Jetson Thor Computers to Advance Mainstream Robotics and Edge AI
General-purpose robots and autonomous machines are moving from research labs to real-world mass-market deployment, creating demand for compact, power-efficient …
GPT-Red: Unlocking Self-Improvement for Robustness
Explore GPT-Red, OpenAI’s automated red teaming system that uses self-play to improve AI safety, alignment, and prompt injection robustness.…

Measuring Balance of Payments Deficits
Balance-of-payments (BOP) deficits have occupied the attention of policymakers and economists for decades, and recent developments have brought renewed interest…
IDC: Why the right networking approach is foundational to agentic AI
Editor’s note: Today we hear from IDC on the results of its 2026 AI in Networking Special Report Survey exploring the enterprises' concerns about networking inf…
How AI is reshaping the future of the AEC industry
AI is poised to rewire the architecture, engineering, and construction sector. Firms that move quickly to reimagine workflows, improve data use, and automate wo…
How physician CEOs hone the double-edged sword of clinical training
The journey from physician to CEO has no prescribed path, but those who’ve reached the top reveal how they evolved the strengths and instincts forged in clinica…

Google named a Leader in the 2026 IDC MarketScape for Worldwide Foundation Model Software
For years, we’ve built with a clear priority: putting the practical needs of the enterprise first. Long before generative AI dominated the headlines, we were fo…
Defense at Machine Speed: The Emerging Architecture Powering AI-Native Cybersecurity
Over the last six months, it seems like every security-oriented conversation has distilled down into one core question: What happens to defense when attackers a…

ScienceSoft’s HIPAA-compliant AI voice scheduler built on AWS
In this post, you will learn how ScienceSoft, an Amazon Web Services (AWS) Services Partner, integrated Amazon Nova 2 Sonic with Amazon Bedrock Guardrails to bu…

Building Service Topology at Scale: Architecture, Challenges, and Lessons Learned

Launching UI for generative AI inference recommendations in Amazon SageMaker AI
In this post, we introduce the UI for optimized generative AI inference recommendations in Amazon SageMaker AI Studio, a low-code no-code (LCNC) experience. The…
Climate planning has prioritized floods. Heat demands equal attention
Summer in the Northern Hemisphere has gotten off to a very hot start. Europe has baked under record-breaking temperatures, while a “heat dome” blankets much of …

Fine-tune NVIDIA Nemotron 3 models with Amazon SageMaker AI serverless model customization
In this post, we explore what makes the Nemotron 3 architecture unique, walk through the fine-tuning techniques available, and show you step-by-step how to get …

Investing in Europe’s digital future: Study projects long-term economic impact of the European Competitiveness Fund
Investing in Europe’s digital future: Study projects long-term economic impact of the European Competitiveness Fund Anonymous (not verified) Thu, 07/09/2026 - 1…
AI is becoming a first hire for small businesses
New research shows how 4 million Americans use ChatGPT to start, run, and grow small businesses, lowering the cost of entrepreneurship with AI.…

Solve harder problems with AlphaEvolve, now available to everyone on Google Cloud
Many of the most challenging and valuable problems in the world are related to optimization. Now, AI is now making these problems tractable. If you've ever trie…

OpenAI finds roughly 30 percent of popular AI coding test is broken
OpenAI reviewed SWE-Bench Pro, a widely used test for measuring AI models' programming skills, and found roughly 30 percent of its tasks are broken. The company…
Separating signal from noise in coding evaluations
A new analysis from OpenAI reveals issues in SWE-Bench Pro, a popular coding benchmark, raising concerns about reliability and accuracy in evaluating AI models.…

ChatGPT can now listen and talk at the same time, making AI conversations seem more human
OpenAI's GPT-Live can listen and speak at the same time using a full-duplex architecture. Complex questions get handed off to GPT-5.5 in the background, which d…

Grok 4.5 is so cheap compared to Fable 5 and GPT 5.5 that benchmark gaps may not matter much
xAI releases Grok 4.5, trained on tens of thousands of Nvidia GB300 GPUs. In coding benchmarks, the model trails Fable 5 and GPT-5.5 but needs 4.2 times fewer t…

Mistral enters robotics with Robostral Navigate, an 8B model that steers robots using just one camera
Mistral is entering the robotics market with Robostral Navigate, an 8B model that guides robots through unknown environments using only a single RGB camera. Tra…

Building and connecting a production-ready ecommerce MCP server using Amazon Bedrock AgentCore and Mistral AI Studio
In this post, you build and connect that server end to end. You will implement MCP tools, set up two-layer JSON Web Token (JWT) authentication, deploy with AWS …