KO EN

AX BRIEF

AI news, benchmarks & engineering blog curation

● LIVE
지능1. Claude Fable 5.1 (max with fallback) 100 pt2. Claude Fable 5 (with fallback) 96 pt3. Claude Opus 5 (max) 95 pt4. GPT-6 Astra (max) 95 pt5. GPT-5.6 Sol (max) 91 pt코딩1. Claude Fable 5.1 (max) 100 pt2. Claude Opus 5 (xhigh) 94 pt3. GPT-6 Astra (max) 91 pt4. Muse Spark 1.3 (xhigh) 81 pt5. Grok 4.5 (high) 81 pt이미지1. GPT Image 2 (high) 100 pt2. MAI-Image-2.6 93 pt3. GPT Image 1.5 (high) 90 pt4. Reve 2.1 88 pt5. Muse Image 85 pt비디오1. Gemini Omni Flash 100 pt2. Minimax H3 Max (post-trained by fal) 94 pt3. Dreamina Seedance 2.0 720p 91 pt4. gemini-omni-1.1-flash 85 pt5. Wan 3.0 85 pt가격1. Claude Fable 5.1 (max with fallback) 100 pt2. GPT-6 Astra (max) 100 pt3. Claude Fable 5 (with fallback) 100 pt4. Claude Opus 5 (max) 49 pt5. GPT-5.6 Sol (max) 39 pt속도1. Gemini 3.5 Flash-Lite 100 pt2. Muse Spark 1.3 (max) 46 pt3. gpt-oss-120b (high) 39 pt4. GPT-5.6 Luna (max) 27 pt5. GPT-5.6 Terra (max) 22 pt

AI News

Daily AI industry news — funding, products, policy, and major moves from global AI companies and startups.

The AI Token Broker Market Trading $100,000 in Daily Inference Credits

The AI Token Broker Market Trading $100,000 in Daily Inference Credits

Token brokers are reselling unused AI inference credits via proxy servers. Some platforms offer 40% discounts and daily spending limits of $

The Trust Crisis Dario Amodei Says Is Fueling AI Backlash

The Trust Crisis Dario Amodei Says Is Fueling AI Backlash

Dario Amodei attributes AI backlash to a systemic crisis of trust. Anthropic proposes regulations that limit frontier AI while aiding smalle

LittleLearner Proves Pre-training Sets the AI Capability Ceiling

LittleLearner Proves Pre-training Sets the AI Capability Ceiling

LittleLearner is an LLM trained exclusively on US K-5 education data. Post-training via GRPO improves in-scope skills but fails to unlock hi

Graph Engineering: The New Orchestration Layer for Probabilistic AI Agents

Graph Engineering: The New Orchestration Layer for Probabilistic AI Agents

AI agent design is evolving from simple loops to complex graph orchestration. Probabilistic LLM nodes require stricter control than traditio

The LLM Eval Harness Exposing Confident Hallucinations in Enterprise AI

The LLM Eval Harness Exposing Confident Hallucinations in Enterprise AI

Arun Mishra developed an eval harness to detect data migration drift. The system found models are most wrong when they are most confident. Q

Why Cloudflare is Pivoting Its Massive Infrastructure Toward AI Agents

Why Cloudflare is Pivoting Its Massive Infrastructure Toward AI Agents

Cloudflare is expanding its AI ecosystem with the Agents SDK and AI Search. The company leverages its massive web traffic to deploy AI agent

AI Math Performance: The Triumph of Working Memory Over Reasoning

AI Math Performance: The Triumph of Working Memory Over Reasoning

AI mathematical success stems from expanded working memory rather than innate intelligence. The model uses its context window as a reasoning

Claude Adopts SynthID-Text to Meet EU AI Act Transparency Standards

Claude Adopts SynthID-Text to Meet EU AI Act Transparency Standards

Anthropic is integrating Google DeepMind's SynthID-Text into Claude. The move ensures compliance with the EU AI Act transparency rules. A ne

Why AI Collaboration is Shifting from Commands to Intent-Centric Leadership

Why AI Collaboration is Shifting from Commands to Intent-Centric Leadership

AI interaction is shifting from deterministic commands to probabilistic intent. Effective collaboration requires structured context and clea

SpaceX Just Bought Cursor: A Massive Power-Up for AI Coding

SpaceX Just Bought Cursor: A Massive Power-Up for AI Coding

SpaceX officially acquires AI code editor Cursor for $60 billion. Cursor now gains access to SpaceX's world-class GPU infrastructure to scal

Ballast for Claude Code: The Plugin That Systematizes AI Workflows

Ballast for Claude Code: The Plugin That Systematizes AI Workflows

Ballast introduces a rule-based system for Claude Code plugins. The tool allows developers to pin verified procedures as permanent rules. It

Suno Studio 2.0 Shifts From AI Generator to Full-Scale DAW

Suno Studio 2.0 Shifts From AI Generator to Full-Scale DAW

Suno Studio 2.0 transforms from a simple generator into a full DAW. The tool introduces MIDI editing and 12-track stem separation. Studio Ch

Claude Code Introduces Loop Engineering via /goal and /loop

Claude Code Introduces Loop Engineering via /goal and /loop

Claude Code introduces loop engineering via /goal and /loop commands. The system separates generation from evaluation to prevent bias. Human

hubble.md Turns Markdown Notes Into Collaborative Agent Workspaces

hubble.md Turns Markdown Notes Into Collaborative Agent Workspaces

hubble.md is an open-source note app designed for human-AI collaboration. The platform supports Markdown and HTML with live reload for agent

DeepSeek V4-Pro Adds Reasoning Control and Off-Peak API Pricing

DeepSeek V4-Pro Adds Reasoning Control and Off-Peak API Pricing

DeepSeek released V4-Pro with adjustable reasoning effort levels. The new API pricing offers 50% discounts during off-peak hours. Native sup

GLM-5.3 Hits 84.5% on CyberGym to Challenge GPT-5.6 Sol

GLM-5.3 Hits 84.5% on CyberGym to Challenge GPT-5.6 Sol

Z.ai released GLM-5.3 focusing on expanded post-training for coding. The model leads in vulnerability discovery but trails in exploit creati

How Claude Code Uses Prompt Caching to Optimize Token Costs

How Claude Code Uses Prompt Caching to Optimize Token Costs

Claude Code utilizes prompt caching to reduce input token costs to 10%. The system prevents context bloat via sub-agents and output limits.

The 3,000 TPS Benchmark Driving Kog's GPU Optimization

The 3,000 TPS Benchmark Driving Kog's GPU Optimization

Kog developed an inference engine achieving 3,000 tokens per second. The software optimizes standard NVIDIA and AMD data center GPUs. The co

Why Google HEIR is the Missing Link for Private AI Inference

Why Google HEIR is the Missing Link for Private AI Inference

Google released HEIR as an open-source compiler for private AI. The tool enables AI inference on data that remains encrypted. Hardware partn

Why Did Qwen Ship 3.8-27B With a Reasoning Effort Parameter?

Why Did Qwen Ship 3.8-27B With a Reasoning Effort Parameter?

Qwen released the 3.8-27B model with controllable reasoning depth. The model supports a native vision-language architecture and 262,144 toke

The $10 Gas Price Risk Facing Meta, Google, and Amazon AI Hubs

The $10 Gas Price Risk Facing Meta, Google, and Amazon AI Hubs

Big Tech is building gigawatt-scale natural gas plants for AI. New pipelines may push regional gas prices above 10 dollars. Rising energy co

Hexis Brings GitOps to AI Agent Skill and Permission Management

Hexis Brings GitOps to AI Agent Skill and Permission Management

Hexis is an open-source control plane for managing AI agent skills. The system uses Git repositories to version control agent permissions. I

The $20 Billion Collapse of Leopold Aschenbrenner's Situational Awareness

The $20 Billion Collapse of Leopold Aschenbrenner's Situational Awareness

Leopold Aschenbrenner's Situational Awareness fund collapsed after heavy losses. AI lab arrogance led to failures in both financial markets

Microsoft Copilot Unifies Consumer and Business Apps to Fight ChatGPT

Microsoft Copilot Unifies Consumer and Business Apps to Fight ChatGPT

Microsoft is merging its consumer and business Copilot apps into a single experience. Several underperforming AI features will be removed by

The Hidden Token Tax of AI Coding Sub-agent Delegation

The Hidden Token Tax of AI Coding Sub-agent Delegation

AI coding sub-agents incur significant fixed token costs during delegation. Prompt cache loss forces sub-agents to re-read project files fro

Why Google Cut Gemini 3.7 Flash API Prices by 50%

Why Google Cut Gemini 3.7 Flash API Prices by 50%

Google released Gemini 3.7 Flash with a 50% API price reduction. The model prioritizes planning to improve first-pass agent accuracy. Benchm

NVIDIA's $500 Billion Plan to Turn GPUs Into Long-Term Infrastructure

NVIDIA's $500 Billion Plan to Turn GPUs Into Long-Term Infrastructure

NVIDIA partnered with global firms for a $500 billion AI fund. The company will guarantee up to 25% of GPU residual values. This shift turns

DeepSeek-V4-Pro-0813 Hits 62.7 Score to Outpace Closed Models

DeepSeek-V4-Pro-0813 Hits 62.7 Score to Outpace Closed Models

DeepSeek-V4-Pro-0813 reaches a 62.7 score on software engineering benchmarks. The model outperforms Opus-4.8 on the Hard LLM Evaluation benc

Capital One Builds Multi-Agent AI Platform Using Custom Llama Models

Capital One Builds Multi-Agent AI Platform Using Custom Llama Models

Capital One developed MACAW to orchestrate custom open-weight models. The platform uses multi-agent workflows to automate complex banking ta

How Google and Anthropic Watermark Gemini and Claude via Token Probability

How Google and Anthropic Watermark Gemini and Claude via Token Probability

Google and Anthropic are integrating watermarks into Gemini and Claude. The system biases token selection probabilities to leave invisible t