KO EN

AX BRIEF

AI news, benchmarks & engineering blog curation

● LIVE
지능1. Claude Fable 5.1 (max with fallback) 100 pt2. Claude Fable 5 (with fallback) 96 pt3. Claude Opus 5 (max) 95 pt4. GPT-6 Astra (max) 95 pt5. GPT-5.6 Sol (max) 91 pt코딩1. Claude Fable 5.1 (max) 100 pt2. Claude Opus 5 (xhigh) 94 pt3. GPT-6 Astra (max) 91 pt4. Muse Spark 1.3 (xhigh) 81 pt5. Grok 4.5 (high) 81 pt이미지1. GPT Image 2 (high) 100 pt2. MAI-Image-2.6 93 pt3. GPT Image 1.5 (high) 91 pt4. Reve 2.1 88 pt5. Muse Image 85 pt비디오1. Gemini Omni Flash 100 pt2. Minimax H3 Max (post-trained by fal) 94 pt3. Dreamina Seedance 2.0 720p 91 pt4. Wan 3.0 85 pt5. gemini-omni-1.1-flash 85 pt가격1. Claude Fable 5.1 (max with fallback) 100 pt2. GPT-6 Astra (max) 100 pt3. Claude Fable 5 (with fallback) 100 pt4. Claude Opus 5 (max) 49 pt5. GPT-5.6 Sol (max) 39 pt속도1. Gemini 3.5 Flash-Lite 100 pt2. Gemini 3.8 Flash (high) 66 pt3. Muse Spark 1.3 (max) 46 pt4. gpt-oss-120b (high) 42 pt5. GPT-5.6 Luna (max) 27 pt

AI Engineering Blogs

Curated deep insights from tech leaders and researchers — engineering, product, and strategy.

Why KIRO is Betting on Physical AI to Hit the 1 Million Robot Goal

Why KIRO is Betting on Physical AI to Hit the 1 Million Robot Goal

KIRO is leading South Korea's push to deploy one million robots by 2030. The institute focuses on localizing core components to reduce forei

IntBot and Certis Shift Humanoid Focus From Manipulation to Interaction

IntBot and Certis Shift Humanoid Focus From Manipulation to Interaction

IntBot and Certis are deploying socially intelligent humanoids in Singapore. The partnership shifts the focus from physical dexterity to soc

DeepAgentTrove Streamlines SFT with 1.7 Million Agent Trajectories

AgentTrove Streamlines SFT with 1.7 Million Agent Trajectories

AgentTrove provides 1.7 million agent interaction trajectories for SFT. The platform uses streaming access to eliminate massive local downlo

Robotis AI Sapiens: Why Open-Sourcing Hardware is the New AI Playbook

Robotis AI Sapiens: Why Open-Sourcing Hardware is the New AI Playbook

Robotis unveiled AI Sapiens, a humanoid capable of text-to-motion control. The robot uses 100% localized parts including DYNAMEXEL-Q actuato

Step 3.7 Flash Hits 97% of Claude Opus 4.6 Performance at $0.19

Step 3.7 Flash Hits 97% of Claude Opus 4.6 Performance at $0.19

StepFun released Step 3.7 Flash featuring a 198B MoE architecture. The model achieves 97% of Claude Opus 4.6 performance at $0.19 per task.

Hermes Agent Cuts MCP Tool Overhead by 85% to Boost Accuracy

Hermes Agent Cuts MCP Tool Overhead by 85% to Boost Accuracy

Hermes Agent reduces MCP tool overhead by 85 percent. The system uses BM25 search to load tool schemas on demand. Tool search improved Opus

Genesis World 1.0 Slashes 200-Hour Robot Tests to 30 Minutes

Genesis World 1.0 Slashes 200-Hour Robot Tests to 30 Minutes

Genesis World 1.0 reduces the sim-to-real gap by 45 percent. The system compresses 200 hours of robot testing into 30 minutes. A zero-shot r

Google's Futures Lab Prototyping the Shift Toward AI-Driven Tutoring

Google's Futures Lab Prototyping the Shift Toward AI-Driven Tutoring

Google and the University of Waterloo launched the Futures Lab for AI tutoring. Students created AI tools for Japanese learning and sign lan

DeepAmazon SageMaker Unifies GPU Metrics and LLM Quality in One Dashboard

Amazon SageMaker Unifies GPU Metrics and LLM Quality in One Dashboard

Amazon SageMaker integrates GPU monitoring with LLM quality metrics. The system uses CloudWatch and Grafana to correlate cost and accuracy.

DeepWhy Google AI Studio's Antigravity Agent Enables Vibe Coding

Why Google AI Studio's Antigravity Agent Enables Vibe Coding

Google AI Studio introduced the Antigravity agent to enable Vibe Coding. Non-developers can now build apps using intent and source materials

Boston Children's Hospital Diagnoses 40 Rare Diseases With AI Infrastructure

Boston Children's Hospital Diagnoses 40 Rare Diseases With AI Infrastructure

Boston Children's Hospital deployed a secure Enterprise AI Layer for all staff. The system diagnosed 40 rare diseases and saved 60,000 work

DeepWhy OpenAI Shipped GPT-Rosalind for Global Biodefense

Why OpenAI Shipped GPT-Rosalind for Global Biodefense

OpenAI launched GPT-Rosalind to combat biological threats. The model provides specialized reasoning for pandemic prevention. Access is restr

The 3.82 Point Boost NVIDIA X-Token Brings to Llama-3.2-1B

The 3.82 Point Boost NVIDIA X-Token Brings to Llama-3.2-1B

NVIDIA X-Token enables knowledge distillation across different tokenizers. The method uses a projection matrix to align mismatched vocabular

Braintrust Erases the Backlog With GPT-5.5 Powered Codex

Braintrust Erases the Backlog With GPT-5.5 Powered Codex

Braintrust adopted GPT-5.5 powered Codex to accelerate development. Half the engineering team switched to the tool within one month. The tea

DeepClaude Opus 4.8: The 99.8% Pass Rate That Changes AI Coding

Claude Opus 4.8: The 99.8% Pass Rate That Changes AI Coding

Anthropic released Claude Opus 4.8 featuring a new Dynamic Workflows system. The model achieved a 99.8% test pass rate on a 750,000-line Rus

DeepThe SageMaker MLflow REST API Proxy Solving Enterprise SDK Bans

The SageMaker MLflow REST API Proxy Solving Enterprise SDK Bans

AWS released a REST API proxy for Amazon SageMaker MLflow. The proxy allows MLflow access without installing the official SDK. This solution

Gemini Omni and 3.5 Flash Shift Google From Search to Execution

Gemini Omni and 3.5 Flash Shift Google From Search to Execution

Google introduced Gemini Omni and 3.5 Flash to enable real-time video editing and agentic workflows. The Antigravity framework allows Gemini

NVIDIA and TorqueAGI Partner to Deploy Physical AI in Enterprise Robots

NVIDIA and TorqueAGI Partner to Deploy Physical AI in Enterprise Robots

NVIDIA and TorqueAGI partnered with John Deere and Dexterity for Physical AI. The collaboration moves AI from digital screens to adaptive in

Google I/O Opens Executive Doors for 10 Changgoo Startups

Google I/O Opens Executive Doors for 10 Changgoo Startups

Google I/O officially invited 10 graduates from the Changgoo program. These startups gained direct access to Google executives and networks.

PyTorch Profiler Reveals the GPU Overhead Killing Your Model's Speed

PyTorch Profiler Reveals the GPU Overhead Killing Your Model's Speed

PyTorch Profiler identifies GPU overhead in NVIDIA A100 environments. Small matrix operations create overhead-bound bottlenecks that waste G

DeepDaegu's AI Humanoid Hub Targets the Manufacturing Labor Gap

Daegu's AI Humanoid Hub Targets the Manufacturing Labor Gap

Daegu launched an AI humanoid hub with a 2.37 billion KRW investment. The center developed a 140cm bipedal platform for factory deployment.

NVIDIA GeForce NOW Leverages RTX 50 Power for 007 First Light Launch

NVIDIA GeForce NOW Leverages RTX 50 Power for 007 First Light Launch

NVIDIA launched 007 First Light on GeForce NOW using RTX 50 GPUs. The game is bundled with a 12-month Ultimate membership subscription. Clou

Endava Scales Senior Engineering Judgment With Codex

Endava Scales Senior Engineering Judgment With Codex

Endava implemented Codex to transform into an Agentic Organization. The system codifies senior engineer judgment to empower junior developer

DeepLFM2.5-8B-A1B Matches 26B Performance Using Only 6GB Memory

LFM2.5-8B-A1B Matches 26B Performance Using Only 6GB Memory

Liquid AI released LFM2.5-8B-A1B for high-performance on-device AI. The model matches Gemma-4-26B performance using only 6GB of memory. New

Why Deep Agent Reliability Requires These 4 Evaluation Patterns

Why Deep Agent Reliability Requires These 4 Evaluation Patterns

Deep agents require trajectory-based evaluation rather than simple input-output checks. A hybrid approach using code, model, and human evalu

Google I/O 2026 Shifts Gemini From Search To Execution

Google I/O 2026 Shifts Gemini From Search To Execution

Google I/O 2026 introduces Gemini Omni and 3.5 Flash for AI agents. Antigravity technology enables real-time generative UI for complex tasks

DeepClaude Opus 4.8 Cuts Code Defects by 4x to Fix AI Overconfidence

Claude Opus 4.8 Cuts Code Defects by 4x to Fix AI Overconfidence

Anthropic released Claude Opus 4.8 with a 4x reduction in code defects. The model introduces adjustable effort levels and dynamic workflows.

NVIDIA's 8 ICRA Frameworks Bridge the Sim-to-Real Robotics Gap

NVIDIA's 8 ICRA Frameworks Bridge the Sim-to-Real Robotics Gap

NVIDIA unveiled eight robotics frameworks at ICRA to bridge the sim-to-real gap. New tools like SPARR and Grasp-MPC significantly increase r

DeepMovensys and Intel Core Ultra Series 3 Eliminate Physical AI Latency

Movensys and Intel Core Ultra Series 3 Eliminate Physical AI Latency

Movensys unveils real-time control technology at the Intel Edge Solution Summit 2026. The system integrates AI decision-making and robot exe

Gemini 3.5 and Intelligent Eyewear Move AI From Screens to Sight

Gemini 3.5 and Intelligent Eyewear Move AI From Screens to Sight

Google unveiled Gemini 3.5 and Intelligent Eyewear at I/O 2026. New tools enable real-time virtual world creation and CG video. AI is shifti