惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Blog — PlanetScale
Blog — PlanetScale
J
Java Code Geeks
月光博客
月光博客
Engineering at Meta
Engineering at Meta
WordPress大学
WordPress大学
Jina AI
Jina AI
小众软件
小众软件
U
Unit 42
云风的 BLOG
云风的 BLOG
Stack Overflow Blog
Stack Overflow Blog
雷峰网
雷峰网
博客园 - Franky
Microsoft Security Blog
Microsoft Security Blog
罗磊的独立博客
宝玉的分享
宝玉的分享
B
Blog
C
Check Point Blog
爱范儿
爱范儿
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
量子位
阮一峰的网络日志
阮一峰的网络日志
Vercel News
Vercel News
酷 壳 – CoolShell
酷 壳 – CoolShell

Recent Commits to openclaw:main

test: merge chat side-result checks · openclaw/openclaw@ddd2c2a test: merge cron history checks · openclaw/openclaw@f7eb746 test: merge responsive navigation shell checks · openclaw/openclaw@c2e4b47 docs(changelog): add codex oauth fixes · openclaw/openclaw@628e6cd test: merge navigation routing cases · openclaw/openclaw@5d8cecb Tests: mock channel registry bundled fallback · openclaw/openclaw@2b08233 Secrets: avoid broad web search discovery for single plugin config · openclaw/openclaw@a464f59 test: merge config view browser checks · openclaw/openclaw@20cf511 fix(status): align oauth health with runtime · openclaw/openclaw@eed7116 feat: add macOS screen snapshots for monitor preview (#67954) thanks … · openclaw/openclaw@f377db1 fix: report shared auth scopes in hello-ok (#67810) thanks @BunsDev · openclaw/openclaw@0b6c39b Auto-reply: avoid eager bundled route fallback · openclaw/openclaw@3ea1bf4 Tests: narrow session binding contract setup · openclaw/openclaw@54e4e16 fix(macOS): enable undo/redo in webchat composer text input (#34962) · openclaw/openclaw@00951dc Tests: speed up channel setup promotion · openclaw/openclaw@82b529a Docs: refresh agent instructions · openclaw/openclaw@5775fe2 fix(auth): serialize OAuth refresh across agents to fix #26322 (#67876) · openclaw/openclaw@8e79080 test: allow ollama public surface boundary test · openclaw/openclaw@7d4f1a6 Docs: add test performance guardrails · openclaw/openclaw@89706d3 Tests: restore context-engine usage proof · openclaw/openclaw@e4c4f95 Tests: slim context engine runtime coverage · openclaw/openclaw@74c198f ci: retry failed custom checkouts · openclaw/openclaw@0ee5baf test: trim duplicate provider auth onboarding cases · openclaw/openclaw@1ffc02e matrix: fix sessions_spawn --thread subagent session spawning (#67643) · openclaw/openclaw@1ce2596 test: reduce auth choice fixture churn · openclaw/openclaw@857b9cd test: mock health status config boundaries · openclaw/openclaw@9d5ab4a test: mock onboard config io boundary · openclaw/openclaw@299694d test: mock legacy state plugin boundaries · openclaw/openclaw@2713089 test: mock channel install boundaries · openclaw/openclaw@b945248 test: mock doctor preview channel boundaries · openclaw/openclaw@b1a3ad4
docs: update OpenAI GPT-5.5 API guidance · openclaw/openc...
steipete · 2026-04-26 · via Recent Commits to openclaw:main

@@ -238,15 +238,15 @@ refs and write a judged Markdown report:

238238239239

```bash

240240

pnpm openclaw qa character-eval \

241-

--model openai/gpt-5.4,thinking=medium,fast \

241+

--model openai/gpt-5.5,thinking=medium,fast \

242242

--model openai/gpt-5.2,thinking=xhigh \

243243

--model openai/gpt-5,thinking=xhigh \

244244

--model anthropic/claude-opus-4-6,thinking=high \

245245

--model anthropic/claude-sonnet-4-6,thinking=high \

246246

--model zai/glm-5.1,thinking=high \

247247

--model moonshot/kimi-k2.5,thinking=high \

248248

--model google/gemini-3.1-pro-preview,thinking=high \

249-

--judge-model openai/gpt-5.4,thinking=xhigh,fast \

249+

--judge-model openai/gpt-5.5,thinking=xhigh,fast \

250250

--judge-model anthropic/claude-opus-4-6,thinking=high \

251251

--blind-judge-models \

252252

--concurrency 16 \

@@ -263,7 +263,7 @@ Use `--blind-judge-models` when comparing providers: the judge prompt still gets

263263

every transcript and run status, but candidate refs are replaced with neutral

264264

labels such as `candidate-01`; the report maps rankings back to real refs after

265265

parsing.

266-

Candidate runs default to `high` thinking, with `medium` for GPT-5.4 and `xhigh`

266+

Candidate runs default to `high` thinking, with `medium` for GPT-5.5 and `xhigh`

267267

for older OpenAI eval refs that support it. Override a specific candidate inline with

268268

`--model provider/model,thinking=<level>`. `--thinking <level>` still sets a

269269

global fallback, and the older `--model-thinking <provider/model=level>` form is

@@ -278,12 +278,12 @@ Candidate and judge model runs both default to concurrency 16. Lower

278278

`--concurrency` or `--judge-concurrency` when provider limits or local gateway

279279

pressure make a run too noisy.

280280

When no candidate `--model` is passed, the character eval defaults to

281-

`openai/gpt-5.4`, `openai/gpt-5.2`, `openai/gpt-5`, `anthropic/claude-opus-4-6`,

281+

`openai/gpt-5.5`, `openai/gpt-5.2`, `openai/gpt-5`, `anthropic/claude-opus-4-6`,

282282

`anthropic/claude-sonnet-4-6`, `zai/glm-5.1`,

283283

`moonshot/kimi-k2.5`, and

284284

`google/gemini-3.1-pro-preview` when no `--model` is passed.

285285

When no `--judge-model` is passed, the judges default to

286-

`openai/gpt-5.4,thinking=xhigh,fast` and

286+

`openai/gpt-5.5,thinking=xhigh,fast` and

287287

`anthropic/claude-opus-4-6,thinking=high`.

288288289289

## Related docs