← Archive

Wednesday, August 19, 2026

27 stories.

01AgentsProduct4 sources agree

Glean introduces Waldo agentic search model

Glean has developed Waldo, a specialized agentic search model that works in concert with frontier models to deliver end-to-end agentic outcomes, reducing latency by 50% and token usage by 25% while maintaining quality. Waldo is trained using a combination of direct preference optimization and reinforcement learning, and is designed to determine how much reasoning a task requires based on its own execution. The model will be deployed to customers soon.

02AgentsProduct4 sources agree

Glean Unveils Third-Generation AI Assistant

Glean introduces its third-generation AI Assistant with advanced personalization and agentic intelligence, and announces significant expansions to its Work AI platform, including new SDKs and MCP capabilities. The new Enterprise Graph underpins these innovations, enabling AI that truly understands the enterprise and its workflows. Glean Assistant now delivers complete outcomes tailored to each employee's way of working, without requiring advanced prompt engineering. The platform also features a new interactive workspace, Canvas, and allows employees to control their Assistant experience. Additionally, Glean is changing the way agents are built and deployed, making it easier for anyone to create and refine agents, and adding richer actions and MCP directory and host support.

03AgentsProduct2 sources agree

Agent Harnesses, Evals, and Production Feedback Loops

Miles v0.1, a new open-source RL framework, has been announced, and multiple developments in agent evaluation, search benchmarking, and harnesses have been reported, including LangSmith Tuned Evaluators and Managed Deep Agents, indicating a shift towards robust rollouts, CI, observability, and environment plumbing. Search benchmarking for agents is maturing, with Artificial Analysis launching its Search Index, and LangChain introducing LangSmith Tuned Evaluators, which claim better performance at lower cost.

04BusinessProduct3 sources agree

Frontier Model Cost and Open-Weights Popularity is Driving Demand for Model Routing

Glean's model routing helps control AI costs for organizations by selecting the most cost-effective model for each task, and its human feedback loop improves the routing system. The company has seen significant growth, reaching $300 million in annual recurring revenue, and is becoming a key player in the enterprise AI market. Glean's architecture includes a model called Waldo, which filters user queries and determines the best model to use, and the company is also seeing increased interest in open-weight models due to cost concerns.

05ModelsProductsingle source

Open Models: Qwen3.8-27B Momentum, GLM-5.3’s Post-Training Gains, and the Small-Model Debate

Qwen3.8-27B has become the top local model in Cline, with impressive benchmark results, while GLM-5.3 has been launched via API with significant gains in intelligence index and Elo rating, driven by stronger post-training techniques. The development of capable local models raises safety implications and highlights a shift towards useful, locally deployable models. GLM-5.3's gains suggest a shift in agentic capability scaling from parameter count to RL systems and environment quality.

06SafetyProductsingle source

OpenAI’s Frontier RL Pause, Expanded Monitoring, and the Shift Toward “Pacing the Frontier”

OpenAI paused some frontier RL training for two weeks to strengthen monitoring, isolation, and red-teaming, and is holding its largest planned frontier RL run, with a focus on hardening security and alignment controls. The slowdown mainly affects farther-out releases, not models already near ship.

07AgentsProductsingle source

Research Notes: Multi-Agent Coordination, Training Variance, and Public AI Usage Measurement

The Public AI Observatory is a new effort to measure real AI assistant usage, with 24,521 consented conversations and 52 models analyzed, while separate research highlights key findings on multi-agent teams and training variance, including the impact of task structure on communication topology and the emergence of specification gaming in agent collectives. The observatory aims to provide public-interest observability for AI usage patterns, independent of vendor reporting.

08On-deviceProduct2 sources agree

Inference and Systems Infra: Mojo Open Source, TensorRT Connect, Cursor’s Git Storage, and Faster Decoding

Modular has open-sourced Mojo, positioning it as a portability layer across accelerators, while NVIDIA launched TensorRT Model Connect for direct model conversion and deployment, and Cursor published a retrospective on Git hosting at scale. Additionally, several companies announced advancements in on-device inference and datacenter accelerators, including DFlash 2 and Cerebras CS-4, highlighting the increasing importance of inference speed.

10CodingProductsingle source

Glean introduces model choice for Assistant conversations

Glean now allows users to select a specific AI model for each Assistant conversation, with options including GPT, Claude, and Gemini models, and also offers an Auto option that automatically selects the best model based on internal evaluations and live usage data. Admins can control which models are available for their users and set default models for their organization.

11AgentsProductsingle source

Top tweets (by engagement)

Anthropic's Claude autonomously designed protein binders for 14 out of 15 targets, and Claude gained Gmail and Google Drive actions, while other AI developments and updates were also announced

16AgentsProduct2 sources agree

Glean outperforms Claude Cowork in cost benchmark

Glean's harness and routing capabilities were benchmarked against Claude Cowork, showing a 4x cost advantage due to lower token volume and cheaper rates. Glean achieved this through model family routing, model tier routing, and better context handling, consuming fewer tokens per query. More results will be presented at Glean:GO!

19AgentsProduct2 sources agree

Zillow adopts Glean for AI-forward culture

Zillow implemented Glean to unify its fragmented data landscape, enabling employees to quickly find information and deploy specialized agents across critical workflows, resulting in improved onboarding, customer focus, and engineering acceleration. Glean's integration with MCP servers and AI coding assistants like Claude Code also accelerated project initiation and raised coding standards.

20BusinessBig picturesingle source

Booking.com adopts Glean AI platform

Booking.com implemented Glean's AI and search platform to improve information access and productivity, resulting in reduced video script creation time and faster IT ticket resolution. The company also integrated AI into its strategy and workflows, adopting Glean as its first company-wide AI platform

From Around the Web

01On-deviceProduct3 sources agree

Cerebras launches CS-4, 30x faster inference than GPUs

Cerebras introduced the CS-4, a rack-scale AI solution that delivers up to 30x faster inference compared to GPUs, with enhanced economics and simplified deployment, featuring three WSE-3 Turbo per system and a modular Nexus Rack-Scale Platform. The CS-4 solution shifts the inference Pareto frontier, delivering up to 10x more throughput per watt than CS-3 while generating tokens up to 30x faster than production GPU systems.

02CodingProduct2 sources agree

The Mojo language (by Modular, now Qualcomm) is now open-source

Modular has made its Mojo language and Modular Cloud platform open source and publicly available, supporting various hardware accelerators including AWS Trainium, Google TPUs, and Qualcomm Cloud AI 100, with native Windows support coming soon. The Modular Platform is now production-ready, serving billions of tokens per minute and powering real enterprise deployments, with flagship customers like MiniMax. Modular is also opening up its MAX licensing model and expanding source access to build an industry alliance program.

03CodingProduct2 sources agree

AI usage patterns in software teams

Linear's report shows a significant increase in AI adoption and pull requests among its users, with coding agents driving most of the acceleration. The report also highlights the blurring of roles, with senior leaders and non-engineers committing code. However, the increased output has not led to time savings, with teams working more, not less.

04CodingProduct2 sources agree

fx releases v0.0.3 coding agent harness

fx is a minimalistic, open-source coding agent harness and CLI written in Zig, optimized for research and embeddability, with a focus on performance and minimalism. It features a small binary size, instant installation, and embedding capabilities, making it suitable for resource-constrained environments and agent sandboxes.

05ResearchInternalssingle source

Palomar: A registry of Lean verified mathematics

The Palomar registry, an initiative for verified mathematics, is now open for submissions of Lean proofs, providing a platform for formalizing and verifying mathematical results. The registry checks submissions for correctness and adherence to best practices, and welcomes both human-generated and AI-generated proofs.

07CodingProduct2 sources agree

Bun 1.4 Rust rewrite is not looking good

The Bun 1.4 rewrite, which heavily utilizes AI-powered tools like Anthropic's Claude, has been plagued by delays and criticism from the community, with over 5,000 open pull requests and concerns about code quality and safety. The project's creator, Jarred, has faced backlash for making repeated promises of imminent release, only to delay again. The rewrite has also raised questions about the effectiveness of AI-assisted coding and the decision to switch from Zig to Rust.