KO EN

AX BRIEF

AI news, benchmarks & engineering blog curation

● LIVE
지능1. Claude Fable 5.1 (max with fallback) 100 pt2. Claude Fable 5 (with fallback) 96 pt3. Claude Opus 5 (max) 95 pt4. GPT-6 Astra (max) 95 pt5. GPT-5.6 Sol (max) 91 pt코딩1. Claude Fable 5.1 (max) 100 pt2. Claude Opus 5 (xhigh) 94 pt3. GPT-6 Astra (max) 91 pt4. Muse Spark 1.3 (xhigh) 81 pt5. Grok 4.5 (high) 81 pt이미지1. GPT Image 2 (high) 100 pt2. MAI-Image-2.6 93 pt3. GPT Image 1.5 (high) 91 pt4. Reve 2.1 88 pt5. Muse Image 85 pt비디오1. Gemini Omni Flash 100 pt2. Minimax H3 Max (post-trained by fal) 94 pt3. Dreamina Seedance 2.0 720p 91 pt4. gemini-omni-1.1-flash 85 pt5. Wan 3.0 85 pt가격1. Claude Fable 5.1 (max with fallback) 100 pt2. GPT-6 Astra (max) 100 pt3. Claude Fable 5 (with fallback) 100 pt4. Claude Opus 5 (max) 49 pt5. GPT-5.6 Sol (max) 39 pt속도1. Gemini 3.5 Flash-Lite 100 pt2. Gemini 3.8 Flash (high) 80 pt3. Muse Spark 1.3 (max) 44 pt4. gpt-oss-120b (high) 40 pt5. GPT-5.6 Luna (max) 27 pt

AI Engineering Blogs

Curated deep insights from tech leaders and researchers — engineering, product, and strategy.

GPT-5.5 Instant Boosts Suicide Risk Detection by 39%

GPT-5.5 Instant Boosts Suicide Risk Detection by 39%

OpenAI updated GPT-5.5 Instant to better detect self-harm risks. A new Safety Summaries system tracks risk signals across conversations. The

DeepHow Amazon Deploys Claude 3.5 Sonnet and RAG for Regulatory Compliance

How Amazon Deploys Claude 3.5 Sonnet and RAG for Regulatory Compliance

Amazon integrates Claude 3.5 Sonnet to automate complex regulatory workflows. The system uses RAG and Amazon Bedrock to query thousands of i

AWS-Cisco Partnership Automates MCP and A2A Agent Security

AWS-Cisco Partnership Automates MCP and A2A Agent Security

AWS and Cisco launched a partnership to automate AI agent security. The system scans MCP servers and A2A agents for prompt injection. Automa

IBM Granite R2 Hits 60.3 MTEB Score With Only 97M Parameters

IBM Granite R2 Hits 60.3 MTEB Score With Only 97M Parameters

IBM released the Granite Embedding Multilingual R2 model series. The 97M parameter model leads its class with a 60.3 MTEB score. A shift to

Amazon Quick and Athena Enable Cross-Account Access for Distributed Costs

Amazon Quick and Athena Enable Cross-Account Access for Distributed Costs

Amazon Quick and Athena now support cross-account access for data analysis. The new role chaining mechanism shifts query costs to the consum

Amazon Bedrock AgentCore Now Controls AI Browsing With 450 Chrome Policies

Amazon Bedrock AgentCore Now Controls AI Browsing With 450 Chrome Policies

Amazon Bedrock AgentCore now supports over 450 Chrome Enterprise policies. The update allows developers to enforce strict URL filtering and

Poetiq Meta-System Pushes GPT 5.5 High to 93.9% Coding Score

Poetiq Meta-System Pushes GPT 5.5 High to 93.9% Coding Score

Poetiq developed a Meta-System that boosts LLM coding performance. GPT 5.5 High reached a 93.9% score on LiveCodeBench Pro. The system impro

Amazon Lex Assisted NLU Reaches 92% Intent Classification Accuracy

Amazon Lex Assisted NLU Reaches 92% Intent Classification Accuracy

Amazon Lex launched Assisted NLU to boost chatbot accuracy using LLMs. The feature achieves 92% intent classification and 84% slot accuracy.

PwC and Anthropic Scale Claude to Cut Underwriting From 10 Weeks to 10 Days

PwC and Anthropic Scale Claude to Cut Underwriting From 10 Weeks to 10 Days

PwC and Anthropic are integrating Claude into finance and healthcare. Insurance underwriting times dropped from 10 weeks to 10 days. PwC is

Why OpenAI Shipped Codex Mobile With Remote SSH Integration

Why OpenAI Shipped Codex Mobile With Remote SSH Integration

OpenAI integrated Codex into the ChatGPT mobile app for remote control. The update enables real-time SSH server management via a secure rela

Eliminating 24.0% GPU Idle Time with Asynchronous Batching

Eliminating 24.0% GPU Idle Time with Asynchronous Batching

GPU idle time accounts for 24.0% of total inference duration. Synchronous batching forces hardware to wait for CPU task preparation. Asynchr

Anthropic and Gates Foundation Launch $200 Million AI Partnership

Anthropic and Gates Foundation Launch $200 Million AI Partnership

Anthropic and the Gates Foundation started a $200 million AI partnership. The collaboration uses Claude to fight diseases and improve global

How Dos Pinos Cut Packaging Errors by 50% Using AI Agents

How Dos Pinos Cut Packaging Errors by 50% Using AI Agents

Dos Pinos reduced packaging design errors by 50% using AI agents. Microsoft Copilot Studio enabled custom agents to verify technical specs.

GeForce NOW Removes Hardware Barriers With Subnautica 2 and Forza Horizon 6

GeForce NOW Removes Hardware Barriers With Subnautica 2 and Forza Horizon 6

NVIDIA GeForce NOW added 11 new games including Subnautica 2. Forza Horizon 6 is now available via early access for premium users. The servi

Pulse AI and Amazon Bedrock Process 1,000 Financial Docs in 3 Hours

Pulse AI and Amazon Bedrock Process 1,000 Financial Docs in 3 Hours

Pulse AI and Amazon Bedrock now process 1,000 financial documents in under 3 hours. The pipeline utilizes the Amazon Nova Micro model with a

AutoScout24 Slashes Development Cycles from Weeks to Days via Codex

AutoScout24 Slashes Development Cycles from Weeks to Days via Codex

AutoScout24 partnered with OpenAI to integrate Codex into its engineering workflow. The company reduced its software development cycles from

How Codex Automates Financial MBRs and Model QA

How Codex Automates Financial MBRs and Model QA

Codex automates the creation of monthly business review narratives. The tool performs automated quality assurance on complex financial model

OpenAI Forces macOS App Updates After TanStack Supply Chain Attack

OpenAI Forces macOS App Updates After TanStack Supply Chain Attack

OpenAI is rotating code signing certificates after a TanStack supply chain attack. Two internal devices were compromised, though no customer

The Codex Financial Workflow Most Finance Teams Haven't Tried Yet

The Codex Financial Workflow Most Finance Teams Haven't Tried Yet

Codex automates the creation of executive financial narratives. The system validates financial models for structural errors and hardcodes. F

Codex Windows Sandbox: Securing AI Agents via SID and Token Isolation

Codex Windows Sandbox: Securing AI Agents via SID and Token Isolation

OpenAI implemented a Windows sandbox for Codex using SIDs and tokens. The system blocks network access without requiring administrator privi

Why AI Chatbots Like Claude Still Leak Private Phone Numbers

Why AI Chatbots Like Claude Still Leak Private Phone Numbers

DeleteMe reports a 400 percent increase in AI-related privacy inquiries. LLMs are leaking personal phone numbers by internalizing training d

Amazon Nova Sonic and WebRTC: Solving Real-Time Voice Latency

Amazon Nova Sonic and WebRTC: Solving Real-Time Voice Latency

Amazon Nova Sonic integrates speech-to-speech processing for AI agents. WebRTC replaces WebSocket to handle unstable network environments ef

Anthropic Launches Claude for Small Business to Automate Daily Workflows

Anthropic Launches Claude for Small Business to Automate Daily Workflows

Anthropic introduced Claude for Small Business to integrate AI into daily tools. The platform supports 15 automated workflows across seven m

Google DeepMind's AI Pointer Understands What You're Pointing At

Google DeepMind's AI Pointer Understands What You're Pointing At

Google DeepMind released a Gemini-based AI mouse pointer this week. The system understands visual and semantic context of cursor targets. Tw

Hermes Agent Hits 140K GitHub Stars, Goes Local on NVIDIA RTX and DGX Spark

Hermes Agent Hits 140K GitHub Stars, Goes Local on NVIDIA RTX and DGX Spark

Hermes agent by Nous Research surpasses 140,000 GitHub stars in three months. It runs entirely locally on NVIDIA RTX PCs, PRO workstations,

NVIDIA and Ineffable Intelligence Co-Design Infrastructure for RL Super-Learners

NVIDIA and Ineffable Intelligence Co-Design Infrastructure for RL Super-Learners

NVIDIA is partnering with Ineffable Intelligence to build RL-specific infrastructure. The collaboration utilizes Grace Blackwell and upcomin

How DeepMind's AlphaFold Nobel Legacy is Powering Google Gemini

How DeepMind's AlphaFold Nobel Legacy is Powering Google Gemini

Google DeepMind evolved from game-playing AI to scientific discovery. Demis Hassabis won the 2024 Nobel Prize for AlphaFold's protein resear

The Hybrid Memory Architecture Powering OpenAI Autonomous Agents

The Hybrid Memory Architecture Powering OpenAI Autonomous Agents

Developers are building autonomous agents with hybrid memory systems. The architecture combines semantic search and BM25 via rank fusion. Mo

AWS P6 Instances and EFA Solve the Foundation Model Bottleneck

AWS P6 Instances and EFA Solve the Foundation Model Bottleneck

AWS introduced P6 instances featuring Blackwell B200 and B300 GPUs. The EFA network reduces latency via OS-bypass RDMA and SRD protocols. Ha

Amazon SageMaker AI Adds FLOPs Meter to Navigate EU AI Act Risks

Amazon SageMaker AI Adds FLOPs Meter to Navigate EU AI Act Risks

Amazon SageMaker AI integrated a FLOPs meter for fine-tuning tracking. The tool helps developers avoid penalties under the EU AI Act 30% rul