AI news, benchmarks & engineering blog curation
Curated deep insights from tech leaders and researchers — engineering, product, and strategy.

OpenAI revealed a security framework for its Codex AI agents. The system uses sandboxing and OpenTelemetry for strict governance. AI-driven

AMD ROCm successfully trained a MedQA model without using CUDA. The project utilized the AMD Instinct MI300X with 192GB of VRAM. High memory

Halliburton integrated Amazon Bedrock into its Seismic Engine software. The AI assistant automates the configuration of 82 specialized tools

OpenAI released three new real-time voice models via API. GPT-Realtime-2 shows a 15.2% improvement in audio intelligence. New pricing models

OpenAI released three specialized models for real-time voice applications. The GPT-Realtime-2 model features a 128K token context window for

Anthropic introduced Natural Language Autoencoders to decode AI activations. The system converts internal numerical states into human-readab

Amazon SageMaker AI implements RLVR and GRPO to stop reward hacking. Qwen2.5-0.5B improves math accuracy using the GSM8K dataset. Rule-based

Google's AlphaEvolve agent reduced genomic analysis errors by 30 percent. The AI optimizes the DeepConsensus model to improve mutation detec

The US Department of Energy and NVIDIA launched the Genesis Mission. The Solstice supercomputer will feature 100,000 Vera Rubin GPUs. AI is

Demis Hassabis praised Korea's manufacturing infrastructure as ideal for AI. Google's Gemini AI is now powering Boston Dynamics' Spot robot

vLLM transitioned from version 0.8.5 to 0.18.1 with major structural changes. Developers found numerical inconsistencies in GSPO reinforceme

KAIST received a Google Foundational Science Grant for geothermal research. The team uses physics-informed AI to predict underground fluid d

Zyphra AI released ZAYA1-8B trained on AMD Instinct MI300x hardware. The model uses 760M active parameters to beat much larger models in mat

Uber integrated OpenAI models to help drivers optimize their daily earnings. A multi-agent architecture and AI Guard ensure low latency and

Parloa launched the AMP platform to manage enterprise voice AI agents. The system uses modular GPT-5.4 agents to reduce human transfers by 8

OpenAI introduced the MRC protocol to eliminate GPU idle time in large clusters. The system uses intelligent packet-spraying to connect 130,

Meta AI released NeuralBench to standardize EEG model evaluation. The framework integrates 94 datasets and 36 distinct brain tasks. Results

OpenAI B2B Signals show leading firms use AI 3.5x more than peers. Cisco saved 1,500 engineering hours monthly using Codex. Enterprise AI su

Anthropic is utilizing SpaceX Colossus 1 resources to expand compute capacity. Claude Code usage limits have doubled for Pro, Max, Team, and

NVIDIA introduced the MRC protocol to Spectrum-X for AI networks. The protocol enables multipath RDMA to reduce GPU idle time. The technolog

The Open ASR Leaderboard added private datasets to prevent benchmark contamination. New evaluation standards use OpenAI Whisper's normalizat

OpenAI released the MRC protocol to the Open Compute Project. The system connects 131,000 GPUs using only two switch layers. MRC reduces tra

Google introduced MTP drafters to accelerate Gemma 4 inference. The technology increases generation speed by up to 3x without quality loss.

OpenAI released a beta Ads Manager for US-based companies to run campaigns. The platform now supports CPC bidding alongside traditional CPM

NVIDIA and ServiceNow unveiled Project Arc for autonomous enterprise tasks. The system leverages Blackwell GPUs to reduce token costs by 35

PORTool introduces a reward tree to optimize AI agent tool calling. The algorithm reduces redundant steps while increasing final accuracy. T

Anthropic released 10 specialized agent templates for financial services. Claude Opus 4.7 achieves 64.37% accuracy on the Vals AI benchmark.

Google introduced event-driven webhooks to the Gemini API. The update eliminates inefficient polling for long-running AI tasks. Developers c

Amazon Quick now integrates directly with S3 Tables. The update removes the need for complex ETL pipelines. Users can analyze Apache Iceberg

OpenAI and PwC are collaborating to build AI agents for corporate finance. OpenAI processed five times more contracts using its own Codex mo