The AI briefing
AI news todayFrom the labs, the papers and the leaderboard.
Original announcements and research, alongside the changes in our rankings. Every headline takes you to its source.
Latest announcements and leaderboard moves
Newest first-
Into the Omniverse: How Developers Turn Ideas Into Simulations With Frontier AI Agents
-
Pay-per-inference for AI agents: How BlockRun and Incarna use Amazon Bedrock AgentCore payments
-
Share GPU clusters across teams with isolation and fairness using Amazon SageMaker HyperPod
-
How Oracle turns days of work into minutes with ChatGPT and Codex
-
Rally Up: ‘Gears of War: E-Day’ Launches on GeForce NOW
-
LegalOn halves Codex costs while maintaining development speed
-
Pollo AI turns creative ideas into campaigns with OpenAI
-
Disrupting AI-enabled “false front” operations
-
The model that didn't exist, so you made it yourself
-
Introducing Claude Haiku 5.5 on AWS
-
NVIDIA, Microsoft Kick Off a New Beginning for Windows PCs With RTX Spark and AI Agents
-
Rethinking access control for RAG with Amazon Quick and Amazon Bedrock
-
Multimodal open d1 decision models for the edge
-
Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses
-
Beyond hours saved: Building the business case for agentic automation
-
How Qlik built grounded, enterprise-scale AI with Amazon Bedrock
-
Automate remediation post AWS DevOps Agent investigation
-
Building AI builders: Playbook for closing the AI knowledge-capability gap
-
How Cornerstone OnDemand cut database diagnosis by 78% with Amazon Bedrock
-
Introducing Falcon ASR
-
One Model Family, Two Gold-Level Results: Fine-Tuning Nemotron for IOI and IMO
-
Helping teens learn, plan, and shape the future of AI
-
Introducing Playground: Create and play custom games
-
Radisson Hotel Group brings hotel discovery into ChatGPT
-
GPT-6 and Intelligent UI for everyone
-
Claude Haiku 5.5 now ranks #64
-
EmbeddingGemma 2: an open, lightweight multimodal embedding model
-
Building a context-aware AI assistant on AgentCore and OpenClaw
-
What AI gets wrong and what failure teaches us
-
Atlassian and OpenAI expand partnership to turn enterprise knowledge into action
-
Responsible AI governance: How AWS positions customers to align with ISO/IEC 42005:2025
-
Best practices for Amazon SageMaker HyperPod administration and governance
-
Manage Amazon SageMaker HyperPod Spaces directly from SageMaker Studio
-
Build a voice travel concierge with Amazon Bedrock AgentCore, Managed Knowledge Base and Nova Sonic
-
Why Telecom Operators Are Building Their AI Strategy on Open Models
-
Introducing Mistral Large 4
-
Sharing AI progress in mathematics
-
How Jump Trading is scaling quant research with ChatGPT
-
Advancing computer use with Ironclad
-
Mistral Large 4 now ranks #53
-
Introducing GLM 5.3 on Amazon Bedrock
-
Supercharge regulated workloads with Claude Code and Amazon Bedrock
-
New agent skill: Amazon SageMaker optimized generative AI inference for your coding agent
-
Making Amazon Quick enterprise-ready: Automated, auditable cross-account resource promotion
-
Agentic retrieval with LangChain and Amazon Bedrock Knowledge Bases
-
Downgrading user roles in Amazon Quick
-
Our approach to EU text provenance rules
-
From Scan to Treatment Plan, AI Helps Close Breast Cancer’s Deadliest Gaps
-
Building advertising for the way people use AI
-
The Agent Said It Was Done. The Database Disagreed.
Recent AI research papers
arXiv · cs.AI / cs.CL / cs.LGPreprints are research claims, not peer-review endorsements. Open each original for the abstract, methods and license.
An Explainable Header-Centric Framework for Large-Scale Semantic Table Interpretation and Data Quality Assessment
Marcelo Valentim Silva, Hannes Herrmann, Valerie Maxville
Teaching PPG How not Who: Fixed-Effects Distillation from ECG
Zhongli Wu, Zhuangzhi Gao, Yuankai Wang, Gregory Y. H. Lip, Bilal H. Kirmani, Yalin Zheng
Beyond the Ergodic Wall: A Discrete Geometric Physics Sandbox for Analysing AI Scaling Limits and Complexity Collapse
Simon Richard Daniel
Clarify, Then Focus: Statement Normalization for Conversation Analytics at Scale
Mikhail L. Arbuzov (Independent researcher), Karan Dave (Independent researcher), Evgeniya Dontsova (Independent researcher), Yaodong Hu (Independent researcher), Vincent Lao (Independent researcher), Navita Jain (Independent researcher), Sisong Bei (Independent researcher), Dmitry Dimov (Independent researcher)
Verification and Self-Improvement in Agentic AI: Foundations and Limits
Chien-Ping Lu
Conversational Task Disambiguation over Tabular Data: Leakage-Aware Formulation, Benchmark Suite, and Training
Nafiseh Ghoroghchian, Luis Scoccola, Tina Sedaghat, Omid Vaheb, Hannah Chen, Dino D'Agostino, Keyvan Golestan
Beyond Owls: Subliminal Learning Can Transfer Learned Capabilities and Backdoors
Jan Dubi\'nski, Anna Sztyber-Betley, Jan Betley, Owain Evans
The Harness as the Only Mutable Surface: Compliance-Bounded Self-Evolution of LLM Agents in Credit Pipelines, with a Measured Admission Gate
Ravil Akhtyamov
Cognitive Thermometers: Machine Learning and Logical Complexity
Shane Steinert-Threlkeld, Jakub Szymanik
Phase-HDC: Replacing Optimizer History with Gradient Thresholds in Discrete Phase Learning
Ahmed Nebli
Coverage, Not Difficulty, Sets How Much Synthetic Data an Activation Probe Needs
Ankush Checkervarty
Synthesis Through Simulation: Generating Coherent Enterprise Data via Scalable Agent-System Interaction
Yipeng Li, Ashutosh Hathidara, Jane Lo, Harshavardhan Abichandani, Gunraj Singh, Atin Ghosh
Visible Reasoning Is Not a Universal Optimizer: Persona- and Thinking-Dependent Effects in Analytics Code Generation
Bhawani Shankar Leelar, Pawan Chorasiya, Davin Hill, Robert E. Tillman, Tamer Soliman
Diffu-LoRA: A Novel Low-Rank Adaptation for Personalized Diffusion Models
Tianjing Li, Wei Zhu
Wieszcz-XIX: A 3.1-Billion-Word Corpus of Pre-1918 Polish and Temporally Bounded Language Models Trained From Scratch
Szymon Kocur
Headlines and publication facts only: no copied article bodies, abstracts or images. Source health and unavailable feeds. “Release-date backfill” items describe current ranking and do not claim a historical rank.