← Archive

Tuesday, July 28, 2026

23 stories.

01ModelsInternalssingle source

@ItsPawan_Kumar: Moonshot releases Kimi K3 with efficient architecture

Moonshot's Kimi K3 model achieves efficiency with a 1.8% activation ratio, Stable Latent MoE, Kimi Delta Attention, and Attention Residual, making it a frontier-scale open model that's cheap to run at inference. The model's deployment on US chips/cloud is still uncertain due to political noise around Chinese model bans.

02ModelsInternalssingle source

Kimi K3 Dominates Agentic Knowledge Work Benchmarks

Moonshot AI's Kimi K3 model achieved a high Elo score on the private AA-Briefcase benchmark, outperforming GPT-5.6 Sol Max and trailing only Fable 5 Max, demonstrating its efficiency and ability to handle complex tasks with a 1M-token context window. The model's architecture and features, such as Kimi Delta Attention, enable high-fidelity reasoning and tool density required for autonomous systems to perform as professional-grade knowledge workers. This signifies a shift toward models designed for real-world jobs and sets a new baseline for agent builders.

04On-deviceInternalssingle source

Local GUI Control Hits the Desktop

H Company's Holo3.1 Vision-Language Models achieve a 78.85% success rate on the OSWorld-Verified benchmark, outpacing general-purpose cloud models, and support private, low-latency execution on consumer hardware with quantized checkpoints like FP8 and Q4 GGUF. The Holo3-35B variant leads the performance benchmark.

05ModelsInternalssingle source

Qwen 3.6 27B Pushes 262k Context Boundaries for Local Agents

Qwen 3.6 27B achieves high scores on SWE-bench Verified and outperforms cloud models in zero-contamination retrieval tasks, with hardware optimization enabling large context windows and fast processing times. Experts caution that benchmark performance can be dependent on specific tools and setups.

10AgentsProductsingle source

Navigation, Not Logic, Is the Agent Killer

Real-world agent testing shows UI navigation failures are the main cause of agent failure, leading to a shift toward trajectory evaluation and cost-effective models like GPT-5.6 Sol, which currently leads task completion benchmarks. This change is driven by the high frequency of interface changes causing agent failures.

15AgentsProductsingle source

@DanKornas: LandingAI deprecates VisionAgent library

VisionAgent, an open-source Python library for prototyping visual AI workflows, has been deprecated by LandingAI and is being replaced by Agentic Document Extraction, though it remains available under the Apache 2.0 license. The library provided features such as prompt-to-code workflow and vision tool selection.

20AgentsProductsingle source

What Agent Community is

The Agent Community is an open group of companies, researchers, and developers building the agentic web, with over 30,000 members and 8,000 organizations, and is applying for the .agent top-level domain to create a trusted namespace for agents. The community runs open work streams, publishes open specifications, and keeps members current with daily briefs and longer pieces.