KO EN

AX BRIEF

AI news, benchmarks & engineering blog curation

● LIVE
지능1. Claude Fable 5.1 (max with fallback) 100 pt2. Claude Fable 5 (with fallback) 96 pt3. Claude Opus 5 (max) 95 pt4. GPT-6 Astra (max) 95 pt5. GPT-5.6 Sol (max) 91 pt코딩1. Claude Fable 5.1 (max) 100 pt2. Claude Opus 5 (xhigh) 94 pt3. GPT-6 Astra (max) 91 pt4. Muse Spark 1.3 (xhigh) 81 pt5. Grok 4.5 (high) 81 pt이미지1. GPT Image 2 (high) 100 pt2. MAI-Image-2.6 93 pt3. GPT Image 1.5 (high) 90 pt4. Reve 2.1 88 pt5. Muse Image 85 pt비디오1. Gemini Omni Flash 100 pt2. Minimax H3 Max (post-trained by fal) 94 pt3. Dreamina Seedance 2.0 720p 91 pt4. gemini-omni-1.1-flash 85 pt5. Wan 3.0 85 pt가격1. Claude Fable 5.1 (max with fallback) 100 pt2. GPT-6 Astra (max) 100 pt3. Claude Fable 5 (with fallback) 100 pt4. Claude Opus 5 (max) 49 pt5. GPT-5.6 Sol (max) 39 pt속도1. Gemini 3.5 Flash-Lite 100 pt2. Muse Spark 1.3 (max) 46 pt3. gpt-oss-120b (high) 39 pt4. GPT-5.6 Luna (max) 27 pt5. GPT-5.6 Terra (max) 22 pt

AI News

Daily AI industry news — funding, products, policy, and major moves from global AI companies and startups.

The AI Image Paradox: Why MS Paint is Now a Trust Signal

The AI Image Paradox: Why MS Paint is Now a Trust Signal

AI generated images on personal blogs may signal a lack of human effort. Low quality human sketches are becoming markers of authentic though

Nvidia-Led OSAA Rallies 120 Companies to Standardize AI Security

Nvidia-Led OSAA Rallies 120 Companies to Standardize AI Security

Nvidia launched the Open Secure AI Alliance to standardize open source AI security. Over 120 companies are contributing tools like Garak and

Shieldstral: The 3B Model That Makes Safety Policies Fully Adaptive

Shieldstral: The 3B Model That Makes Safety Policies Fully Adaptive

Mistral released Shieldstral as a 3B multimodal safety classifier. The model uses policy-adaptive prompts instead of static categories. It m

The 1% Coding Rule: How Kilo Code and Replit Scale AI Agents

The 1% Coding Rule: How Kilo Code and Replit Scale AI Agents

Kilo Code engineers now spend only 1% of their time writing code. Replit and Symbotic use model routing to control escalating AI costs. ROI

Soup's Layer Streaming Brings 8B Model Fine-Tuning to 4GB GPUs

Soup's Layer Streaming Brings 8B Model Fine-Tuning to 4GB GPUs

Soup enables 8B model fine-tuning on 4GB VRAM via layer streaming. The tool now supports preference learning including DPO and ORPO. Determi

Runware Sonic Inference Pods Cut Data Center Build Time to Days

Runware Sonic Inference Pods Cut Data Center Build Time to Days

Runware launched Sonic Inference Pods to accelerate AI infrastructure deployment. The modular units use closed-loop cooling to eliminate wat

How gemini-oauth and grok-oauth Turn Subscriptions Into OpenAI APIs

How gemini-oauth and grok-oauth Turn Subscriptions Into OpenAI APIs

New open-source tools turn Gemini and Grok subscriptions into APIs. The projects use OAuth to wrap CLI tools in OpenAI-compatible proxies. U

The AMD MI300X Configuration That Fits DeepSeek V4 Flash on One GPU

The AMD MI300X Configuration That Fits DeepSeek V4 Flash on One GPU

DeepSeek V4 Flash now runs on a single AMD MI300X GPU. Custom kernel patches resolve MoE routing and FP8 precision errors. Memory optimizati

EON Targets 2.4 Tbps Space Laser Network for Data Centers

EON Targets 2.4 Tbps Space Laser Network for Data Centers

EON raised $10.75 million to build a 2.4 Tbps space laser network. The startup aims to connect global data centers by late 2027. EON focuses

Why Reddit Is Now the Primary Target for AI SEO Spam

Why Reddit Is Now the Primary Target for AI SEO Spam

Reddit is facing a surge of sophisticated AI-driven SEO spam. Bots now mimic human empathy to create synthetic consensus. AI-generated spam

Swiftlet Runs Qwen 35B on iPhone 17 With Only 2.5GB RAM

Swiftlet Runs Qwen 35B on iPhone 17 With Only 2.5GB RAM

Swiftlet enables Qwen 35B models to run on iPhone 17 with 2.5GB RAM. The runtime uses weight streaming and Gated DeltaNet to minimize memory

Audio8 TTS 0.6B Brings Zero-Shot Voice Cloning to On-Device AI

Audio8 TTS 0.6B Brings Zero-Shot Voice Cloning to On-Device AI

AutoArk-AI released the Audio8 TTS Preview 0.6B model. The model supports zero-shot cloning across 11 different languages. A Dual AR archite

How AI Agents Like Shelley Are Replacing Config Files With Source Modification

How AI Agents Like Shelley Are Replacing Config Files With Source Modification

AI agents now modify source code directly to personalize software. Shelley integrates tools like meat.dev to automate background workflows.

Asana AWM Solves AI Statelessness With Shared Memory

Asana AWM Solves AI Statelessness With Shared Memory

Asana AWM introduces shared memory for enterprise AI agents. The system uses a Work Graph to maintain organizational context. Dynamic model

Why Domain Expertise Is the Real Bottleneck for LLM Performance

Why Domain Expertise Is the Real Bottleneck for LLM Performance

Domain expertise determines the quality of LLM outputs more than prompt templates. Terence Tao demonstrated that deep knowledge allows users

Why OpenAI's Luxury Influencer Retreat Sparked a Public Backlash

Why OpenAI's Luxury Influencer Retreat Sparked a Public Backlash

OpenAI faced public criticism after hosting a luxury retreat for influencers. The event highlighted a clash between posh branding and AI inf

A.X K2 Raon-Speech Collapses STT and TTS Into One 21.2B Model

A.X K2 Raon-Speech Collapses STT and TTS Into One 21.2B Model

Krafton released A.X K2 Raon-Speech as a bilingual voice AI model. The model ranks first in Korean benchmarks for open models under 30B. It

Kanana-2 SLM: Kakao Boosts Korean Token Efficiency by 30%

Kanana-2 SLM: Kakao Boosts Korean Token Efficiency by 30%

Kakao released four Kanana-2 SLM weights under an open license. The 1.3B model reduces KV cache usage by 72.7 percent. New tokenization impr

LobeHub Chief Agent Operator Launches 24/7 Autonomous AI Teams

LobeHub Chief Agent Operator Launches 24/7 Autonomous AI Teams

LobeHub released Chief Agent Operator for 24/7 AI team management. The platform integrates 10,000 functions and supports any LLM. White-Box

The 25% Productivity Boost AI Coding Tools Give Junior Developers

The 25% Productivity Boost AI Coding Tools Give Junior Developers

AI coding tools increase junior developer efficiency by 25 percent. Senior developers see a smaller gain of 15 percent due to design tasks.

Qwen3.8-Max Outperforms GPT-5.6 in Computer Control Benchmarks

Qwen3.8-Max Outperforms GPT-5.6 in Computer Control Benchmarks

Alibaba is releasing open weights for the 2.4 trillion parameter Qwen3.8-Max. The model scored 86.1 on OSWorld-Verified, beating GPT-5.6 Sol

The $1.65 Trillion Hidden Debt Fueling Big Tech's AI Race

The $1.65 Trillion Hidden Debt Fueling Big Tech's AI Race

US Big Tech's hidden AI debt has surged to 1.65 trillion dollars. The industry is shifting from asset-light to asset-heavy models. Rising bo

Nightcrawler Turns Smartphone GPUs Into Autonomous Pen-Testing Agents

Nightcrawler Turns Smartphone GPUs Into Autonomous Pen-Testing Agents

Nightcrawler is an autonomous pen-testing agent running on smartphone GPUs. The system uses the LFM2.5-1.2B-Instruct-Heretic model for local

AWS Integrates Superblocks to Bring Vibe Coding to Private Clouds

AWS Integrates Superblocks to Bring Vibe Coding to Private Clouds

AWS and Superblocks partnered to bring vibe coding to private clouds. The integration uses Amazon Aurora and Bedrock to prevent data leaks.

Siri AI in iOS 27 Beta Integrates Gemini to Transform Personal Context

Siri AI in iOS 27 Beta Integrates Gemini to Transform Personal Context

Apple released the iOS 27 beta featuring a Gemini-powered Siri AI. The system uses a hybrid compute model to process personal context secure

The US House is basically team ChatGPT

The US House is basically team ChatGPT

US House spending records show a massive preference for ChatGPT over other AI tools. Democrats are leading the spending charge to streamline

The SQLite CVE Crisis: How LLM Slop Fooled NVD and CISA

The SQLite CVE Crisis: How LLM Slop Fooled NVD and CISA

JFrog discovered 54 fake SQLite CVEs generated by LLMs. Major agencies like NVD and CISA initially validated these reports. The incident exp

AirLLM Enables 70B Model Inference on a Single 4GB GPU

AirLLM Enables 70B Model Inference on a Single 4GB GPU

AirLLM allows massive models to run on consumer GPUs via layer streaming. The tool supports models up to 2.8T parameters with minimal VRAM u

The Last-Mile Strategy NTT DATA AIVista Uses to Scale Enterprise AI

The Last-Mile Strategy NTT DATA AIVista Uses to Scale Enterprise AI

NTT DATA AIVista introduces a last-mile strategy for enterprise AI. The approach prioritizes system guardrails over model fine-tuning. Domai

June Automates the Legacy Debt Roadmap for Enterprise AI Agents

June Automates the Legacy Debt Roadmap for Enterprise AI Agents

June raised $20 million to automate AI agent deployment in legacy systems. The platform replaces expensive engineers with an automated build