KO EN

AX BRIEF

AI news, benchmarks & engineering blog curation

● LIVE
지능1. Claude Fable 5.1 (max with fallback) 100 pt2. Claude Fable 5 (with fallback) 96 pt3. Claude Opus 5 (max) 95 pt4. GPT-6 Astra (max) 95 pt5. GPT-5.6 Sol (max) 91 pt코딩1. Claude Fable 5.1 (max) 100 pt2. Claude Opus 5 (xhigh) 94 pt3. GPT-6 Astra (max) 91 pt4. Muse Spark 1.3 (xhigh) 81 pt5. Grok 4.5 (high) 81 pt이미지1. GPT Image 2 (high) 100 pt2. MAI-Image-2.6 93 pt3. GPT Image 1.5 (high) 91 pt4. Reve 2.1 88 pt5. Muse Image 85 pt비디오1. Gemini Omni Flash 100 pt2. Minimax H3 Max (post-trained by fal) 94 pt3. Dreamina Seedance 2.0 720p 91 pt4. gemini-omni-1.1-flash 85 pt5. Wan 3.0 85 pt가격1. Claude Fable 5.1 (max with fallback) 100 pt2. GPT-6 Astra (max) 100 pt3. Claude Fable 5 (with fallback) 100 pt4. Claude Opus 5 (max) 49 pt5. GPT-5.6 Sol (max) 39 pt속도1. Gemini 3.5 Flash-Lite 100 pt2. Gemini 3.8 Flash (high) 80 pt3. Muse Spark 1.3 (max) 44 pt4. gpt-oss-120b (high) 40 pt5. GPT-5.6 Luna (max) 27 pt

AI News

Daily AI industry news — funding, products, policy, and major moves from global AI companies and startups.

Cursor and the $60 Billion Shift Toward AI Judgment Data

Cursor and the $60 Billion Shift Toward AI Judgment Data

AI agents are causing the collapse of traditional per-seat pricing models. Cursor demonstrates how user correction data creates a durable AI

Show GN Automates Flutter and Flame Game Releases via Claude Code

Show GN Automates Flutter and Flame Game Releases via Claude Code

Show GN is an open-source plugin for Claude Code that automates game publishing. The tool handles everything from initial planning to App St

Brown University's 48-Point Average: The ChatGPT Wake-Up Call

Brown University's 48-Point Average: The ChatGPT Wake-Up Call

Brown University uncovered a massive ChatGPT cheating scandal in an economics course. Student averages plummeted from 96 to 48 when exams mo

GLM 5.2 Beats Claude Code in IDOR Detection with 39% F1 Score

GLM 5.2 Beats Claude Code in IDOR Detection with 39% F1 Score

Zhipu AI released GLM 5.2 as an open-weight MoE model. It outperforms Claude Code in IDOR detection with a 39% F1 score. The MIT license all

Claude Code's Opus 4.8 Challenges a Grade III MRI Tear Diagnosis

Claude Code's Opus 4.8 Challenges a Grade III MRI Tear Diagnosis

A user employed Claude Code's Opus 4.8 to analyze complex MRI DICOM data. The AI contradicted a medical diagnosis of a Grade III partial ten

Prompt Injection: The LLM01 Vulnerability Threatening Enterprise AI

Prompt Injection: The LLM01 Vulnerability Threatening Enterprise AI

OWASP identifies prompt injection as the most critical LLM vulnerability for 2025. Zero-click attacks like EchoLeak can leak enterprise data

Wayfinder Router Eliminates the Cost of LLM Routing Decisions

Wayfinder Router Eliminates the Cost of LLM Routing Decisions

Wayfinder Router directs prompts to local or cloud LLMs using structural analysis. The system eliminates routing costs by avoiding separate

Micron Hits $1.27 Trillion Market Cap Amid AI Memory Shortages

Micron Hits $1.27 Trillion Market Cap Amid AI Memory Shortages

Micron's market cap reached $1.27 trillion due to surging AI memory demand. The company secured 16 strategic long-term supply agreements wit

Princeton's AI RFIC Tool Ends the Era of Manual Black Magic Design

Princeton's AI RFIC Tool Ends the Era of Manual Black Magic Design

Princeton researchers developed an AI system for RFIC design. The tool uses RL and diffusion models to cut design time to minutes. AI-genera

Why Ford Rehired 350 Veteran Engineers After AI Automation Failures

Why Ford Rehired 350 Veteran Engineers After AI Automation Failures

Ford rehired 350 veteran engineers to fix AI automation errors. Automation failures led to billions in losses and high recall rates. Ford re

Why AI Slop Makes Human Experience the New Content Currency

Why AI Slop Makes Human Experience the New Content Currency

AI slop is flooding the internet with lifeless synthetic content. Competitive advantage is shifting from data access to lived experience. Cr

Paca Launches Open-Source Project Management for AI-Driven Scrum Teams

Paca Launches Open-Source Project Management for AI-Driven Scrum Teams

Paca is an open-source project management platform that treats AI agents as team members. The system integrates with Claude via the Model Co

Why SpaceX's Orbital Data Center Vision Clashes With AI's Urgency

Why SpaceX's Orbital Data Center Vision Clashes With AI's Urgency

SpaceX proposes orbital data centers to bypass terrestrial regulations. Critics argue the timeline fails to address the immediate AI compute

Why Claude Code Makes Product Planning the New Engineering Bottleneck

Why Claude Code Makes Product Planning the New Engineering Bottleneck

Anthropic's Claude Code is accelerating development cycles by up to three times. The engineering bottleneck has shifted from implementation

How Claude's 90% Probability Analysis Stopped Unnecessary Radiation

How Claude's 90% Probability Analysis Stopped Unnecessary Radiation

A patient used Claude to synthesize complex data for a rare lymphoma diagnosis. The AI identified a thymus rebound with 90% probability to a

Why Sakana AI and 360 are Filling the Anthropic Cybersecurity Gap

Why Sakana AI and 360 are Filling the Anthropic Cybersecurity Gap

The US government banned non-US access to Anthropic's Mythos and Fable 5. Sakana AI and 360 launched local alternatives to ensure AI soverei

NYT Amends Lawsuit to Target Microsoft's Supercomputing Infrastructure

NYT Amends Lawsuit to Target Microsoft's Supercomputing Infrastructure

The New York Times is amending its lawsuit against Microsoft and OpenAI. NYT claims Microsoft built supercomputers to facilitate copyright t

The 91.91% Terminal-Bench Score Driving GPT-5.6's Sol Model

The 91.91% Terminal-Bench Score Driving GPT-5.6's Sol Model

OpenAI introduced GPT-5.6 with a three-tier system called Sol, Terra, and Luna. The Sol model achieved a 91.91% score on the Terminal-Bench

Samsung's Harness Engineering Automates MCU Firmware Development

Samsung's Harness Engineering Automates MCU Firmware Development

Samsung verified Harness Engineering for autonomous MCU firmware development. The system achieved a 95% completion rate and reduced dev time

OpenTag: The Open-Source Slack Agent Ending Per-Seat AI Pricing

OpenTag: The Open-Source Slack Agent Ending Per-Seat AI Pricing

CopilotKit released OpenTag as an open-source Slack AI agent. The tool replaces per-seat pricing with a self-hosted runtime. It features Gen

Why the US Government Only Allowed 100 Institutions to Use Claude Mythos 5

Why the US Government Only Allowed 100 Institutions to Use Claude Mythos 5

The US government granted 100 trusted institutions access to Claude Mythos 5. This move follows a brief ban on Mythos 5 and Fable 5 due to s

Apple Skips M6 Pro and Max to Fast-Track AI-Centric M7

Apple Skips M6 Pro and Max to Fast-Track AI-Centric M7

Apple will skip M6 Pro and Max chips to prioritize the AI-focused M7. M6 base chips will offer 200GB/s bandwidth to support local AI models.

Why AI Intelligence Now Depends on SMRs and Physical Infrastructure

Why AI Intelligence Now Depends on SMRs and Physical Infrastructure

AI growth is shifting from a software bottleneck to a power and hardware crisis. Big Tech is securing nuclear energy and SMRs to bypass fail

Claude Mythos 5 Returns to 100 US Entities After Security Ban

Claude Mythos 5 Returns to 100 US Entities After Security Ban

US Commerce Secretary Howard Lutnick authorized the return of Claude Mythos 5. Over 100 trusted US government agencies and companies now hav

MRAgent Reduces Long-Term Memory Tokens to 118K per Query

MRAgent Reduces Long-Term Memory Tokens to 118K per Query

MRAgent reduces prompt token usage to 118K per query. The framework uses a Cue-Tag-Content graph for active memory. An automated distillatio

Axonius Warns Autonomous Security AI is Blind to Network Dark Matter

Axonius Warns Autonomous Security AI is Blind to Network Dark Matter

Axonius reveals that 12.7% of network devices lack critical security agents. Autonomous AI agents risk failure by trusting incomplete asset

Weave Router Slashes LLM Costs by 40-70% With Smart Proxying

Weave Router Slashes LLM Costs by 40-70% With Smart Proxying

Weave released a model routing proxy that reduces LLM costs by 40-70%. The system uses a tiny on-box embedder to route requests in under 50m

Tech Influence Watch Exposes AI and Crypto's Million-Dollar Election Play

Tech Influence Watch Exposes AI and Crypto's Million-Dollar Election Play

AI and crypto firms are spending hundreds of millions to influence US elections. Tech Influence Watch tracks these capital flows to expose p

OpenAI Taps Former Uber Chief to Scale India Operations

OpenAI Taps Former Uber Chief to Scale India Operations

OpenAI appointed Prabhjeet Singh as its first Managing Director for India. The company is expanding its physical footprint to New Delhi, Mum

Why the US Government is Approving GPT 5.6 Customer by Customer

Why the US Government is Approving GPT 5.6 Customer by Customer

The US government now controls the release of AI models like GPT 5.6. Anthropic models Fable and Mythos were recalled by federal authorities