{"version":1,"type":"story","url":"https://digestai.news/story/anthropic-and-openai-adopt-chinese-kv-cache-tricks-slash-cache-read-pr","json":"https://digestai.news/story/anthropic-and-openai-adopt-chinese-kv-cache-tricks-slash-cache-read-pr.json","markdown":"https://digestai.news/story/anthropic-and-openai-adopt-chinese-kv-cache-tricks-slash-cache-read-pr.md","slug":"anthropic-and-openai-adopt-chinese-kv-cache-tricks-slash-cache-read-pr","headline":"Anthropic and OpenAI adopt Chinese KV cache tricks, slash cache-read pricing","summary":"Western AI companies are now using a set of key‑value cache optimizations that Chinese lab DeepSeek released publicly.\n\nAnthropic’s Claude Opus 5.5 and OpenAI’s GPT‑6.1 Sol have incorporated these techniques, cutting their cache‑read pricing by 60 % and 80 % respectively versus the prior versions. The lower memory footprint reduces VRAM costs for long‑context inference, making top‑tier models more margin‑positive for the Western providers.\n\nThe article notes that Chinese labs have been openly sharing such performance breakthroughs, a shift from earlier “distillation” concerns, and suggests the move may help Western firms stay competitive despite earlier GPU access constraints.","keyPoints":["Claude Opus 5.5 cuts cache‑read pricing by 60% and GPT‑6.1 Sol by 80% after using the Chinese optimizations","Smaller cache lowers VRAM needs for long‑context models, improving inference margins for Western AI firms"],"whyItMatters":"Reducing KV cache size and pricing makes long‑context AI models cheaper to run, boosting profitability for Western providers and expanding affordable access to advanced AI services.","category":{"slug":"models","name":"Generative AI & Models","url":"https://digestai.news/category/models"},"entities":{"companies":["DeepSeek","Anthropic","OpenAI"],"models":["Claude Opus 5.5","GPT-6.1 Sol","DeepSeek-V4.1-Flash"],"people":[]},"firstPublishedAt":"2026-09-30T15:50:11Z","updatedAt":"2026-09-30T15:50:11Z","sourceCount":1,"hasPrimarySource":false,"sources":[{"outlet":"insufferable.dev","title":"The AI Race Just Got Awkward","url":"https://insufferable.dev/posts/the-ai-race-just-got-awkward","publishedAt":"2026-09-30T15:50:11Z","type":"press","primary":false,"lead":true}],"sourceNotes":null,"discussions":[{"site":"Hacker News","url":"https://news.ycombinator.com/item?id=49910553","points":105}],"thread":{"title":"China's Open Source AI Conquest","url":"https://digestai.news/thread/z-ai-deploys-glm5-3flash-on-100-000chip-chinese-cluster","storyCount":4},"cite":{"text":"Digest AI, \"Anthropic and OpenAI adopt Chinese KV cache tricks, slash cache-read pricing\", 30 September 2026, https://digestai.news/story/anthropic-and-openai-adopt-chinese-kv-cache-tricks-slash-cache-read-pr","publisher":"Digest AI","title":"Anthropic and OpenAI adopt Chinese KV cache tricks, slash cache-read pricing","datePublished":"2026-09-30T15:50:11Z","url":"https://digestai.news/story/anthropic-and-openai-adopt-chinese-kv-cache-tricks-slash-cache-read-pr"},"generatedBy":"Written by Digest AI's editorial model from the linked sources; the sources are the record.","license":"Headlines, digests and key points are written by Digest AI and may be quoted with a link to the story page. Linked articles belong to their publishers. Terms: https://digestai.news/terms#reuse"}