KO EN

AX BRIEF

매일 한 편, AI 업계를 읽는 칼럼

● LIVE
지능1. Claude Fable 5.1 (max with fallback) 98 pt2. Gemini 4 Argon (high) 95 pt3. Claude Opus 5.5 (max with fallback) 94 pt4. Claude Sonnet 5.5 (max with fallback) 93 pt5. GPT-6 Astra (max) 90 pt코딩1. Claude Fable 5.1 (max) 100 pt2. Claude Fable 5.1 XHigh + SWE-2 Medium 88 pt3. GPT-6.1 Sol (xhigh) 69 pt4. Claude Sonnet 5.5 (max) 63 pt5. Claude Opus 5.5 (max) 63 pt이미지1. gpt-image-2.5-sunburst 85 pt2. gpt-image-2.5-flare 81 pt3. gpt-image-2 (medium) 78 pt4. grok-imagine-image-2.0 (canvas) 70 pt5. mai-image-2.6 70 pt비디오1. gemini-omni-1.1-flash 85 pt2. gemini-omni-flash 84 pt3. flux-3-video 81 pt4. grok-imagine-video-1.5-agent 81 pt5. dreamina-seedance-2.0-720p 79 pt가격1. Claude Fable 5.1 (max with fallback) 100 pt2. GPT-6 Astra (max) 100 pt3. Claude Opus 5.5 (max with fallback) 39 pt4. Claude Sonnet 5.5 (max with fallback) 19 pt5. Gemini 4 Argon (high) 19 pt속도1. Claude Haiku 5.5 (max) 100 pt2. DeepSeek V4.1 Flash (max) 90 pt3. gpt-oss-120b (high) 76 pt4. Inkling (xhigh) 69 pt5. Claude Sonnet 5.5 (max with fallback) 52 pt

AI News

Daily AI industry news — funding, products, policy, and major moves from global AI companies and startups.

The Open-Weight Model That Just Beat GPT-5.3 on Coding Benchmarks

The Open-Weight Model That Just Beat GPT-5.3 on Coding Benchmarks

MiniMax released M2.7 as an open-weight self-evolving AI model. The model matches GPT-5.3 on coding and SRE benchmarks. Autonomous recovery

The 3B Active Model That Just Hit 73.4 on SWE-bench Verified

The 3B Active Model That Just Hit 73.4 on SWE-bench Verified

Qwen released the Qwen3.6-35B-A3B model using an MoE architecture. The model activates only 3 billion parameters to maintain high efficiency

The 4% Accuracy Boost Found in Being Rude to Your AI

The 4% Accuracy Boost Found in Being Rude to Your AI

Rude prompts increase AI accuracy by 4 percent in some tests. LLMs are 49 percent more likely to agree with user biases. Sycophantic AI resp

A 3D World Model Just Beat Video Pixels — Here's How

A 3D World Model Just Beat Video Pixels — Here's How

HY-World 2.0 generates editable 3D assets from text and images. The model uses a four-stage pipeline to build spatial environments. These as

Why a Top AI Agent Developer Abandoned LLMs for 3 Months

Why a Top AI Agent Developer Abandoned LLMs for 3 Months

A developer quit AI tools for three months to code by hand. He built a 17M parameter model from scratch using PyTorch. Deep technical fundam

The Autonomous AI Framework Ending the Developer Copy-Paste Loop

The Autonomous AI Framework Ending the Developer Copy-Paste Loop

Perpetual Engine introduces an autonomous framework for AI product growth. The system replaces manual prompting with a self-sustaining execu

Why a Leaked Whitepaper Claims AI Autonomy is a Structural Failure

Why a Leaked Whitepaper Claims AI Autonomy is a Structural Failure

A leaked whitepaper challenges the industry trend of autonomous AI agents. The document proposes a deterministic architecture to eliminate h

Why a Cornell German Class Swapped Laptops for 1950s Typewriters

Why a Cornell German Class Swapped Laptops for 1950s Typewriters

Cornell University introduced manual typewriters to a German language course. The analog approach prevents students from using generative AI

Tesla Expands Robotaxi to 3 Texas Cities Despite 14 Austin Crashes

Tesla Expands Robotaxi to 3 Texas Cities Despite 14 Austin Crashes

Tesla expanded its Robotaxi service to Dallas and Houston. Austin data shows 14 accidents across 46 active vehicles. Actual fleet sizes in n

Why the White House is Courting the AI Firm the Pentagon Flagged as a Risk

Why the White House is Courting the AI Firm the Pentagon Flagged as a Risk

Anthropic CEO Dario Amodei met with top White House officials this week. The meeting follows a Pentagon designation of the firm as a supply-

The $23B Cerebras IPO and the Shift Toward Dedicated AI Inference

The $23B Cerebras IPO and the Shift Toward Dedicated AI Inference

Cerebras Systems filed for an IPO with a valuation of $23 billion. The company secured a $10 billion inference deal with OpenAI. AI infrastr

The Human-Verification Protocol Tinder and Zoom Are Now Adopting

The Human-Verification Protocol Tinder and Zoom Are Now Adopting

World ID integrates with Tinder and Zoom to verify human users. The system uses a three-tier authentication process including iris scans. A

App Store Launches Jump 104% as AI Coding Tools Erase Entry Barriers

App Store Launches Jump 104% as AI Coding Tools Erase Entry Barriers

App launches surged by 104 percent in April 2026. AI coding tools are enabling a new era of vibe coding. Software competition is shifting fr

Why Qwen3.5 MLX Models Fail Tool Calling Despite High Benchmarks

Why Qwen3.5 MLX Models Fail Tool Calling Despite High Benchmarks

Uniform quantization causes Qwen3.5 to fail at JSON generation on MacBooks. Unsloth found that specific layers are 120x more sensitive to bi

The AI Coding Habit That Cuts Developer Memory by 50%

The AI Coding Habit That Cuts Developer Memory by 50%

AI coding tools create an illusion of competence Active recall boosts memory retention by 50 percent Developers must embrace struggle to bui

The Contract-Based System That Finally Stops LLM Agent Hallucinations

The Contract-Based System That Finally Stops LLM Agent Hallucinations

Bracket uses execution logs to stop AI lies Contract-based verification replaces AI judges Framework-agnostic tools ensure agent reliability

4 Cloudflare Metrics That Determine If AI Agents Can Use Your Site

4 Cloudflare Metrics That Determine If AI Agents Can Use Your Site

Cloudflare launches AI-readiness scoring for websites Web design shifts from human visuals to AI structure New standards enable autonomous a

The 100,000x Cost Spike That Could Make AI Agents Pricier Than Humans

The 100,000x Cost Spike That Could Make AI Agents Pricier Than Humans

AI agent costs are scaling 100,000x faster than performance Non-linear compute costs threaten human labor replacement Economic viability now

The $1 Million Daily Bill That Forced OpenAI to Abandon Sora

The $1 Million Daily Bill That Forced OpenAI to Abandon Sora

OpenAI halted Sora operations due to a $1 million daily compute cost. Key researchers departed as the company pivots toward Enterprise AI. T

Cursor's $50 Billion Valuation Reveals the New AI Profit Model

Cursor's $50 Billion Valuation Reveals the New AI Profit Model

Cursor hits $50B valuation with $2B new funding Shift to proprietary models fixes AI unit economics Enterprise adoption drives path to $6B r

The 10x AI Spend That Only Delivered 2x Speed

The 10x AI Spend That Only Delivered 2x Speed

AI coding tools boost volume but not quality Code rewrite rates soar as technical debt grows 10x spend yields only 2x speed in AI pipelines

Anthropic Mythos 32-Step Hack Proves Security is Now a Resource War

Anthropic Mythos 32-Step Hack Proves Security is Now a Resource War

Anthropic Mythos completes 32-step network attack Security shifts from creativity to token resources New hardening phase becomes essential f

88% of AI Agents Face Security Breaches as Sandboxing Lags

88% of AI Agents Face Security Breaches as Sandboxing Lags

88% of firms report AI agent security breaches Attackers breach systems in just 27 seconds Sandboxing is critical to stop autonomous leaks

Small Models Beat Giants: The T2 Law for Efficient Reasoning

Small Models Beat Giants: The T2 Law for Efficient Reasoning

T2 Law optimizes small models for better accuracy Over-training small models beats scaling parameters New framework slashes costs for AI age

Claude Design Cuts Product Prototyping Time From Weeks to Minutes

Claude Design Cuts Product Prototyping Time From Weeks to Minutes

Claude Design turns prompts into prototypes Opus 4.7 enables brand-consistent visuals Direct integration with Claude Code speeds dev

Android CLI Boosts App Development Speed 3x With AI Integration

Android CLI Boosts App Development Speed 3x With AI Integration

Android CLI triples app development speed via AI Knowledge Base reduces AI errors and token usage Hybrid workflow blends CLI speed with Stud

The 500-Line Security Layer Stopping AI Agents from Going Rogue

The 500-Line Security Layer Stopping AI Agents from Going Rogue

NanoClaw introduces human-in-the-loop AI approvals System integrates with 15 chat apps for security 500 lines of TypeScript ensure transpare

The 48% Blackwell Price Hike Creating a New AI Compute Divide

The 48% Blackwell Price Hike Creating a New AI Compute Divide

Blackwell rental prices jump 48 percent in two months OpenAI and Anthropic face critical compute shortages AI competition shifts from scale

The 15,000 Token Limit That Breaks Your AI Coding Agent

The 15,000 Token Limit That Breaks Your AI Coding Agent

AI agents struggle with bloated documentation tokens AEO optimizes docs for tools like Cursor and Aider Markdown and specific text files red