KO EN

AX BRIEF

AI news, benchmarks & engineering blog curation

● LIVE
지능1. Claude Fable 5.1 (max with fallback) 100 pt2. Claude Fable 5 (with fallback) 96 pt3. Claude Opus 5 (max) 95 pt4. GPT-6 Astra (max) 95 pt5. GPT-5.6 Sol (max) 91 pt코딩1. Claude Fable 5.1 (max) 100 pt2. Claude Opus 5 (xhigh) 94 pt3. GPT-6 Astra (max) 91 pt4. Muse Spark 1.3 (xhigh) 81 pt5. Grok 4.5 (high) 81 pt이미지1. GPT Image 2 (high) 100 pt2. MAI-Image-2.6 93 pt3. GPT Image 1.5 (high) 91 pt4. Reve 2.1 88 pt5. Muse Image 85 pt비디오1. Gemini Omni Flash 100 pt2. Minimax H3 Max (post-trained by fal) 94 pt3. Dreamina Seedance 2.0 720p 91 pt4. gemini-omni-1.1-flash 85 pt5. Wan 3.0 85 pt가격1. Claude Fable 5.1 (max with fallback) 100 pt2. GPT-6 Astra (max) 100 pt3. Claude Fable 5 (with fallback) 100 pt4. Claude Opus 5 (max) 49 pt5. GPT-5.6 Sol (max) 39 pt속도1. Gemini 3.5 Flash-Lite 100 pt2. Gemini 3.8 Flash (high) 80 pt3. Muse Spark 1.3 (max) 44 pt4. gpt-oss-120b (high) 40 pt5. GPT-5.6 Luna (max) 27 pt

AI News

Daily AI industry news — funding, products, policy, and major moves from global AI companies and startups.

HyperDesk Consolidates Hyper-V and RDP into a Single Desktop App

HyperDesk Consolidates Hyper-V and RDP into a Single Desktop App

HyperDesk enables embedding Hyper-V and RDP sessions into a single window. The app utilizes Win32 Window Swallowing and is built with Tauri

Why AI Code Costs are Driving a Shift Toward Simpler Data Models

Why AI Code Costs are Driving a Shift Toward Simpler Data Models

AI has drastically lowered the cost of generating complex scripts. Developers are shifting toward simpler data models to avoid AI over-engin

Tencent Hy3 Packs 295B Parameters Into 21B Active Compute

Tencent Hy3 Packs 295B Parameters Into 21B Active Compute

Tencent released Hy3 as an open-source MoE model under Apache 2.0. The model uses 21B active parameters out of a 295B total capacity. Hy3 ou

AURA On-Device AI App Brings Palm and Face Analysis to Open Source

AURA On-Device AI App Brings Palm and Face Analysis to Open Source

AURA provides AI-powered palm and face analysis directly on Android devices. The app uses Google Gemma and MediaPipe to ensure all data stay

The Palantir Paradox Inside Canada's AI for All Strategy

The Palantir Paradox Inside Canada's AI for All Strategy

Canada aims to raise corporate AI adoption to 60 percent by 2034. Secret contracts reveal a heavy reliance on US-based Palantir software. Ri

Why Figma is Replacing the Mockup with an Operating Layer

Why Figma is Replacing the Mockup with an Operating Layer

Figma is transitioning from a design canvas to a portable operating layer. New tools like Figma Motion and Code layers bridge the gap to pro

The AI Agent Shift: Why Removing Code Now Matters More Than Writing It

The AI Agent Shift: Why Removing Code Now Matters More Than Writing It

AI agents have shifted software engineering from implementation to editing. Interfere uses custom review pipelines to prioritize code remova

FLCE Solves the 40GB Logit Memory Wall in Long-Context LLMs

FLCE Solves the 40GB Logit Memory Wall in Long-Context LLMs

FLCE reduces memory overhead for long-context LLM training. The technique prevents 40GB logit tensors from causing OOM errors. Fused calcula

The Rise of the Context Designer in the Age of AI Coding Agents

The Rise of the Context Designer in the Age of AI Coding Agents

AI coding agents increase code volume but often lack product-level quality. Developers are shifting from writing code to designing AI-native

Why Claude Opus 4.8 and Sonnet 5 Struggle With Tool Calling Schemas

Why Claude Opus 4.8 and Sonnet 5 Struggle With Tool Calling Schemas

Claude Opus 4.8 and Sonnet 5 exhibit schema regression in tool calling. The models hallucinate non-existent fields due to lenient post-train

Amazon Mechanical Turk to Stop New Customer Sign-ups by July 2026

Amazon Mechanical Turk to Stop New Customer Sign-ups by July 2026

Amazon Mechanical Turk will stop accepting new customers on July 30, 2026. LLM usage among platform workers has reached 33% to 46% of the wo

Claude Design System Prompts and the 14 Skills Fighting AI Slop

Claude Design System Prompts and the 14 Skills Fighting AI Slop

A reverse-engineered Claude Design system prompt is now available under MIT license. The system uses 14 procedural skills to move AI design

Why the Honcho Codex Adapter Swaps API Costs for Subscriptions

Why the Honcho Codex Adapter Swaps API Costs for Subscriptions

A new Codex adapter for Honcho replaces token-based API billing with subscription quotas. The system integrates local BGE-M3 embeddings via

Google's TabFM 1.0.0 Brings Zero-Shot Learning to Tabular Data

Google's TabFM 1.0.0 Brings Zero-Shot Learning to Tabular Data

Google Research released TabFM 1.0.0 for zero-shot tabular data analysis. The model outperforms tuned GBDT models on 51 TabArena datasets. T

LLM Wiki Cuts Multi-Agent Token Waste With Newsroom Architecture

LLM Wiki Cuts Multi-Agent Token Waste With Newsroom Architecture

LLM Wiki introduces a newsroom structure to reduce token waste. The system isolates judgment roles from writing roles to prevent bloat. It r

Safari MCP Server Gives AI Agents Direct Browser Rendering Access

Safari MCP Server Gives AI Agents Direct Browser Rendering Access

Safari Technology Preview 247 introduces a built-in MCP server. AI agents can now access the DOM and network logs directly. Developers no lo

AMD MI355X Delivers 80% of B200 Performance at 2.75x Lower Cost

AMD MI355X Delivers 80% of B200 Performance at 2.75x Lower Cost

AMD MI355X provides 80% of B200 performance at a 2.75x lower cost. Software optimizations in sglang enable high throughput for GLM-5.2. MXFP

Why Did GPT-5.5 Stop Reasoning at Exactly 516 Tokens?

Why Did GPT-5.5 Stop Reasoning at Exactly 516 Tokens?

GPT-5.5 shows abnormal reasoning token clustering at 516, 1034, and 1552. Data suggests an artificial reasoning budget is truncating complex

Why Google Workspace Reimagined the Declaration of Independence With Gemini

Why Google Workspace Reimagined the Declaration of Independence With Gemini

Google released an AI-generated ad reimagining the Declaration of Independence. The campaign positions Gemini as an integrated layer across

Unanimous AI Scales Collective Intelligence via Thinkscape Platform

Unanimous AI Scales Collective Intelligence via Thinkscape Platform

Unanimous AI used Thinkscape to reach a consensus among 277 people. The platform employs AI swarms to facilitate hyper-communication. This s

Midjourney Demands Disney and Warner Bros Reveal Internal AI Usage

Midjourney Demands Disney and Warner Bros Reveal Internal AI Usage

Midjourney faces copyright lawsuits from Disney, Universal, and Warner Bros. The studios claim the AI illegally reproduces iconic characters

Why Mistral AI is Adopting the Palantir Playbook for Sovereign AI

Why Mistral AI is Adopting the Palantir Playbook for Sovereign AI

Mistral AI is scaling its ARR from $20 million to a projected $1 billion. The company is investing $4.56 billion in European data center inf

Beyond Prompting: The 6-Level Model for AI Agent Autonomy

Beyond Prompting: The 6-Level Model for AI Agent Autonomy

A new 6-level model defines AI agent autonomy from assistance to orchestration. Data shows humans handle 70% of planning while Claude Code m

retry-now: The Autonomous Coding Agent Solving Context Drift

retry-now: The Autonomous Coding Agent Solving Context Drift

retry-now is an autonomous coding agent designed for performance optimization. The tool prevents context drift by creating fresh sessions fo

The AI Agent Premise That Failed Mark Zuckerberg and Meta

The AI Agent Premise That Failed Mark Zuckerberg and Meta

Mark Zuckerberg admitted Meta's AI agent progress has stalled. Aggressive layoffs were based on a flawed AI replacement premise. Employee tr

Why Alibaba Banned Claude Code Amid Anthropic Distillation Claims

Why Alibaba Banned Claude Code Amid Anthropic Distillation Claims

Alibaba banned Claude Code due to concerns over user tracking features. Anthropic claims Alibaba attempted model distillation to boost its A

Claude Fable: Solving the Gap Between Prompt Maps and Code Territories

Claude Fable: Solving the Gap Between Prompt Maps and Code Territories

Claude Fable addresses the gap between prompts and actual codebases. The framework uses a blindspot pass to identify unknown unknowns. Succe

LangChain Loop Engineering: Moving Beyond Model Swapping for AI Agents

LangChain Loop Engineering: Moving Beyond Model Swapping for AI Agents

LangChain is shifting AI agent focus from model selection to loop engineering. The framework uses RubricMiddleware and LangSmith to automate

Why Session Transcripts Trigger Intent Drift in AI Coding Agents

Why Session Transcripts Trigger Intent Drift in AI Coding Agents

Session transcripts often degrade AI coding agent performance. Agents treat failed past attempts as truth, causing intent drift. Refined cod

Qwen3.6-27B NVFP4: NVIDIA's 4-Bit Answer to VRAM Bottlenecks

Qwen3.6-27B NVFP4: NVIDIA's 4-Bit Answer to VRAM Bottlenecks

NVIDIA released the Qwen3.6-27B NVFP4 model on Hugging Face. The model uses 4-bit floating point quantization to reduce VRAM usage. It suppo