惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

博客园 - 三生石上(FineUI控件)
月光博客
月光博客
人人都是产品经理
人人都是产品经理
Google DeepMind News
Google DeepMind News
M
MIT News - Artificial intelligence
Vercel News
Vercel News
MyScale Blog
MyScale Blog
爱范儿
爱范儿
博客园 - 司徒正美
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
IT之家
IT之家
H
Help Net Security
Last Week in AI
Last Week in AI
阮一峰的网络日志
阮一峰的网络日志
酷 壳 – CoolShell
酷 壳 – CoolShell
L
LangChain Blog
罗磊的独立博客
Stack Overflow Blog
Stack Overflow Blog
宝玉的分享
宝玉的分享
博客园 - 聂微东
云风的 BLOG
云风的 BLOG
J
Java Code Geeks
博客园 - 叶小钗
D
Docker

Release Notes on DigitalOcean Documentation

Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note Release Note
Release Note
DigitalOcean · 2026-07-01 · via Release Notes on DigitalOcean Documentation

Last verified 1 Jul 2026

Prompt caching for open-source models in serverless inference chat completions and responses API is now in public preview. Open-source models cache context automatically, so you do not need to set the cache_control or prompt_cache_retention parameters.

Prompt caching is available for the following open-source models:

  • DeepSeek V3.2
  • DeepSeek V4 Pro
  • DeepSeek V4 Flash
  • Kimi K2.5
  • Kimi K2.6
  • GLM 5
  • GLM-5.1
  • GLM-5.2
  • gpt-oss-120b
  • MiMo V2.5
  • MiMo V2.5 Pro
  • MiniMax M2.5
  • Qwen 3.5
  • Qwen3 Coder Flash

For more information, see Use Prompt Caching.