Caching
Caching on Beyond Market Intelligence: a running collection of 5 stories we have gathered and hand-picked because they are worth your time. Every post here touches on caching in some way — the news, the analysis, the deep dives, and the occasional surprise find. Acme AI is the next-generation, AI-powered spreadsheet platform built to replace Excel and redefine how analysts, data scientists, and enterprise teams work with data. New stories are added to this page as we find them, so check back if you want to keep up with what is happening around caching, or subscribe to the RSS feed to get them as soon as they are published. Browse the collection below, or head back to the homepage to see everything Beyond Market Intelligence is covering right now.

Anthropic's Claude Fable 5.1 and Mythos 5.1 arrive with a 75% cost reduction for Fable cache reads
Anthropic has released Claude Fable 5.1 and Claude Mythos 5.1, the latest iterations of its powerful large language models, alongside a significant 75% cost reduction for Fable cache reads. These models prioritize sustained problem-solving, demonstrating substantial improvements on benchmarks like Terminal-Bench and AutomationBench. Crucially, Anthropic is also introducing Enterprise Frontier Safeguards (EFS), allowing organizations to retain monitoring data within their own infrastructure. This release addresses evolving enterprise needs for capable, economical, and governable AI agents—a shift underscored by recent cybersecurity evaluations.

GLM-5.3 hits the API at $1.4/$4.4 per million tokens
Z.ai has made GLM-5.3, its new open-source language model boasting advanced coding and agent capabilities, accessible via API. Developers can now integrate this frontier model into their applications at a competitive rate of $1.40 per million input tokens and $4.40 per million output tokens—unchanged from its predecessor, GLM-5.2. Independent benchmarks place GLM-5.3 among the world’s top open-weight models, demonstrating strong performance at a notably lower cost than premium alternatives. For teams exploring coding and agent workloads, GLM-5.3 represents a compelling, accessible option.

DeepSeek's top-ranked V4 Flash stumbles on real agent tasks as its prices surge
DeepSeek’s V4 Flash, initially lauded as a "total monster" for its impressive leaderboard performance and remarkably low pricing, is experiencing a shift in perception. Recent testing reveals it completes only 53.8% of complex agent tasks in real-world scenarios. Simultaneously, DeepSeek is adjusting its pricing model, increasing rates by as much as 1,100% for certain token types.

Cloudflare Introduces Cache Response Rules for Post-Origin Cache Control
Cloudflare's latest innovation, Cache Response Rules, represents a significant advancement in post-origin cache control. Previously limited to request attributes, Cache Rules now evaluate origin responses *before* they’re cached, providing granular control over what content enters the Cloudflare network. This rules engine empowers developers to optimize caching strategies and improve performance with unprecedented precision. For deeper insights into Cloudflare’s ongoing enhancements, explore "Cloudflare Adds Agent Tracing," detailing new span capabilities for agent invocations.

DoorDash Uses Envoy and Valkey for a 1.5M RPS Proxy Cache with 99.99999% Availability
DoorDash achieves unparalleled data efficiency with Entity Cache, a novel proxy caching platform built on Envoy and Valkey. This innovative solution reduces redundant service-to-service requests within their microservices architecture, handling over 1.5 million requests per second with an impressive 99.99999% availability. Through caching, event-driven invalidation, and robust failure handling, Entity Cache optimizes performance and ensures consistent reliability. For those interested in exploring related advancements in data analysis, consider our survey on deep learning for scRNA-seq analysis.