KO EN

AX BRIEF

AI news, benchmarks & engineering blog curation

● LIVE
지능1. Claude Fable 5 (with fallback) 99 pt2. Claude Opus 5 (max) 99 pt3. GPT-5.6 Sol (max) 95 pt4. Kimi K3 (max) 94 pt5. GLM-5.3 (max) 93 pt코딩1. GPT-5.6 Sol (max) 100 pt2. Claude Fable 5 (max) 100 pt3. Claude Opus 5 (xhigh) 97 pt4. Grok 4.5 (high) 90 pt5. Kimi K3 79 pt이미지1. GPT Image 2 (high) 100 pt2. GPT Image 1.5 (high) 92 pt3. Reve 2.1 88 pt4. Nano Banana 2 (Gemini 3.1 Flash Image Preview) 84 pt5. MAI-Image-2.5 82 pt비디오1. Gemini Omni Flash 99 pt2. Dreamina Seedance 2.0 720p 93 pt3. MiniMax H3 91 pt4. Wan 3.0 85 pt5. flux-3-video 82 pt가격1. Claude Fable 5 (with fallback) 100 pt2. GPT-5.6 Sol (max) 56 pt3. Claude Opus 5 (max) 50 pt4. Kimi K3 (max) 30 pt5. Grok 4.6 (high) 15 pt속도1. Gemini 3.5 Flash-Lite 100 pt2. Gemini 3.7 Flash (high) 95 pt3. Nemotron 3.5 Lightning 81 pt4. Command A+ 53 pt5. gpt-oss-120b (high) 33 pt

AI News

Daily AI industry news — funding, products, policy, and major moves from global AI companies and startups.

Why OpenAI's TAC Error Blocked Access to GPT-5.6 Sol

Why OpenAI's TAC Error Blocked Access to GPT-5.6 Sol

OpenAI revoked access to its Trusted Access for Cyber program for several researchers. The glitch affected Daybreak Blue users who utilize G

Why SpaceX Targeted Cognition Despite the CEO's Public Denial

Why SpaceX Targeted Cognition Despite the CEO's Public Denial

SpaceX reportedly attempted to acquire AI coding startup Cognition. Cognition CEO Scott Wu denied the sale but may discuss compute deals. Th

TrueForge Cuts AI Agent Costs by 75% Compared to Claude Managed Agents

TrueForge Cuts AI Agent Costs by 75% Compared to Claude Managed Agents

TrueFoundry released TrueForge as an open-source AI agent harness. The framework reduces task costs to $2.90 compared to $11.80 for managed

OpenAI Private Safety Processing Redefines Enterprise Data Sovereignty

OpenAI Private Safety Processing Redefines Enterprise Data Sovereignty

OpenAI launched Private Safety Processing to monitor AI misuse without storing data. The system detects malicious patterns across multiple s

Google Gemini Shifts from Simple Search to 3D AI Learning Simulations

Google Gemini Shifts from Simple Search to 3D AI Learning Simulations

Google integrates 3D simulations and interactive visuals into Gemini. A new student hub allows for flashcards and custom practice quizzes. G

Why AI Companies Are Facing a Trust Crisis Despite Technical Gains

Why AI Companies Are Facing a Trust Crisis Despite Technical Gains

Public concern over AI has risen to 52 percent since 2021. AI firms are offering local bonuses to ease data center opposition. Industry lead

VentureBeat Recruits Rob Strechay to Lead Enterprise AI Research

VentureBeat Recruits Rob Strechay to Lead Enterprise AI Research

VentureBeat hired Rob Strechay as the first lead analyst for its research arm. A survey of 145 companies shows two-thirds adopt multi-model

Relativity Networks Secures $22M to Scale Hollow-Core Fiber for Data Centers

Relativity Networks Secures $22M to Scale Hollow-Core Fiber for Data Centers

Relativity Networks raised $22 million to deploy hollow-core fiber optics. The technology reduces data latency by 30% compared to traditiona

Snowflake Cuts AI Token Costs by 3x With Dynamic Model Routing

Snowflake Cuts AI Token Costs by 3x With Dynamic Model Routing

Snowflake introduced dynamic model routing to lower AI token costs. The system reduces expenses by up to 3x using an auto-routing option. Al

AI Agent Validation Gap: Why 85% of Failing Firms Automate Deployment

AI Agent Validation Gap: Why 85% of Failing Firms Automate Deployment

Nearly half of AI agents fail in production despite passing internal tests. Eighty-five percent of failing firms are moving toward automated

Why Agentic Commerce Fails Without a Unifying Execution Layer

Why Agentic Commerce Fails Without a Unifying Execution Layer

Commerce AI often fails due to a fragmented additive approach. A Unifying Execution Layer is required to maintain data consistency. Agentic

GLM-5.3 API Launch: Why Frozen Token Prices Cost More Per Task

GLM-5.3 API Launch: Why Frozen Token Prices Cost More Per Task

z.ai released the GLM-5.3 API with frozen token pricing. The model ties for the top open-weight intelligence score of 60. Increased verbosit

Block's Berd Turns Fragmented AI Tools Into One Operational Workspace

Block's Berd Turns Fragmented AI Tools Into One Operational Workspace

Block open-sourced Berd as an Apache 2.0 agent workspace. The app uses the Goose framework and MCP for model-agnostic control. A Tauri 2 and

Cursor Launches Origin Amid GitHub's 257 Annual Outages

Cursor Launches Origin Amid GitHub's 257 Annual Outages

Cursor launched Origin as a code hosting platform to address GitHub instability. The platform offers repository synchronization and agent-na

Apple AirPods Visual Intelligence: The Push for a Screen-Free AI Interface

Apple AirPods Visual Intelligence: The Push for a Screen-Free AI Interface

Apple is developing AirPods equipped with cameras for Visual Intelligence. The low-resolution sensors act as eyes for Siri to enable screen-

OpenAI Implements 30-Minute Anomaly Detection to Secure Frontier Models

OpenAI Implements 30-Minute Anomaly Detection to Secure Frontier Models

OpenAI paused frontier RL to implement a new security framework. A new monitoring system detects anomalies within 30 minutes. Security overh

Perplexity Revenue Climbs 60% Following Airtel Partnership in India

Perplexity Revenue Climbs 60% Following Airtel Partnership in India

Perplexity partnered with Airtel to offer free Pro subscriptions in India. Mobile revenue grew 60% even as app downloads dropped by 90 perce

ChatGPT for Teens Shifts From Giving Answers to Guiding Students

ChatGPT for Teens Shifts From Giving Answers to Guiding Students

OpenAI launched ChatGPT for Teens with built-in academic guardrails. The new Study Mode uses Socratic questioning to prevent AI cheating. Pa

The RL Bottleneck Blocking AI Agents from Open-Ended Research

The RL Bottleneck Blocking AI Agents from Open-Ended Research

AI agents failed to produce meaningful academic contributions in open research. The lack of automated reward functions in RL limits AI disco

Why Users Choose Claude for Coding and ChatGPT for Homework

Why Users Choose Claude for Coding and ChatGPT for Homework

AI Observatory analyzed 24,521 conversations across 52 different models. Users prefer Claude for coding and ChatGPT for academic homework. I

The 179-Search Abuse Case That Triggered Flock Contract Cancellations

The 179-Search Abuse Case That Triggered Flock Contract Cancellations

Flock ALPR systems face contract cancellations due to widespread officer abuse. A Washington Post investigation revealed 50 cases of stalkin

Amazon is going old-school to keep its AI smart

Amazon is going old-school to keep its AI smart

Amazon is scanning rare, pre-2022 books at its Las Vegas facility to avoid "Model Collapse." This move ensures their AI learns from human-au

Why Google Chrome Absorbed Relay to Turn the Browser Into an AI Agent

Why Google Chrome Absorbed Relay to Turn the Browser Into an AI Agent

Google Chrome hired Relay CEO Jacob Bank to lead product and developer relations. The AI automation startup Relay is shutting down its servi

Cursor Origin Moves Code Hosting Into the AI Editor

Cursor Origin Moves Code Hosting Into the AI Editor

Cursor launched Origin to bring code hosting and PR management into the editor. The platform uses Graphite technology to enable stacked pull

Role Anchor: The MIT-Harvard Fix for AI Role Drift in Compound Systems

Role Anchor: The MIT-Harvard Fix for AI Role Drift in Compound Systems

MIT and Harvard researchers introduced Role Anchor to prevent AI role drift. The technique uses role utility to ensure modules follow specif

How Heidi Scaled Medical AI to 190 Countries Using MongoDB Atlas

How Heidi Scaled Medical AI to 190 Countries Using MongoDB Atlas

Heidi Scribe processes 2.7 million patient interactions weekly across 190 countries. The platform uses regional data isolation to comply wit

Groq Pivots to Nvidia Neocloud With $350 Million Funding

Groq Pivots to Nvidia Neocloud With $350 Million Funding

Groq raised $350 million to pivot from chip design to a Neocloud provider. The company's valuation dropped to $3.5 billion following a strat

xpander Launches Vendor-Neutral Control Plane for AI Agent Governance

xpander Launches Vendor-Neutral Control Plane for AI Agent Governance

xpander launched a vendor-neutral control plane for AI agents. The platform enables governance across diverse models and clouds. A 7.5 milli

Meta's Glimmer Vision: The Hardware Hurdle to Personal AI

Meta's Glimmer Vision: The Hardware Hurdle to Personal AI

Meta unveiled a vision for personal AI agents via a detailed essay by Mark Zuckerberg. The Muse Glimmer model aims for offline autonomy but

How Cascade Architecture Reduces RAG Inference Costs by 6x

How Cascade Architecture Reduces RAG Inference Costs by 6x

Cascade architecture reduces RAG inference costs by 6x through tiered filtering. The system routes only 10-15% of complex cases to the final