惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

S
Security Affairs
S
Schneier on Security
N
News | PayPal Newsroom
T
Threatpost
Cloudbric
Cloudbric
H
Heimdal Security Blog
Recent Commits to openclaw:main
Recent Commits to openclaw:main
Google Online Security Blog
Google Online Security Blog
D
Darknet – Hacking Tools, Hacker News & Cyber Security
Spread Privacy
Spread Privacy
V
Vulnerabilities – Threatpost
The Last Watchdog
The Last Watchdog
CTFtime.org: upcoming CTF events
CTFtime.org: upcoming CTF events
L
LINUX DO - 最新话题
P
Proofpoint News Feed
C
CXSECURITY Database RSS Feed - CXSecurity.com
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
Apple Machine Learning Research
Apple Machine Learning Research
NISL@THU
NISL@THU
Application and Cybersecurity Blog
Application and Cybersecurity Blog
The Hacker News
The Hacker News
O
OpenAI News
人人都是产品经理
人人都是产品经理
C
Cyber Attacks, Cyber Crime and Cyber Security
C
Check Point Blog
C
Cisco Blogs
GbyAI
GbyAI
J
Java Code Geeks
L
LangChain Blog
I
Intezer
T
Tailwind CSS Blog
有赞技术团队
有赞技术团队
MyScale Blog
MyScale Blog
美团技术团队
The Register - Security
The Register - Security
Help Net Security
Help Net Security
WordPress大学
WordPress大学
Y
Y Combinator Blog
T
Tor Project blog
M
MIT News - Artificial intelligence
爱范儿
爱范儿
TaoSecurity Blog
TaoSecurity Blog
V
Visual Studio Blog
T
Threat Research - Cisco Blogs
P
Palo Alto Networks Blog
月光博客
月光博客
T
Tenable Blog
S
Securelist
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
D
DataBreaches.Net

Recent Commits to openclaw:main

test: merge chat side-result checks · openclaw/openclaw@ddd2c2a test: merge cron history checks · openclaw/openclaw@f7eb746 test: merge responsive navigation shell checks · openclaw/openclaw@c2e4b47 docs(changelog): add codex oauth fixes · openclaw/openclaw@628e6cd test: merge navigation routing cases · openclaw/openclaw@5d8cecb Tests: mock channel registry bundled fallback · openclaw/openclaw@2b08233 Secrets: avoid broad web search discovery for single plugin config · openclaw/openclaw@a464f59 test: merge config view browser checks · openclaw/openclaw@20cf511 fix(status): align oauth health with runtime · openclaw/openclaw@eed7116 feat: add macOS screen snapshots for monitor preview (#67954) thanks … · openclaw/openclaw@f377db1 fix: report shared auth scopes in hello-ok (#67810) thanks @BunsDev · openclaw/openclaw@0b6c39b Auto-reply: avoid eager bundled route fallback · openclaw/openclaw@3ea1bf4 Tests: narrow session binding contract setup · openclaw/openclaw@54e4e16 fix(macOS): enable undo/redo in webchat composer text input (#34962) · openclaw/openclaw@00951dc Tests: speed up channel setup promotion · openclaw/openclaw@82b529a Docs: refresh agent instructions · openclaw/openclaw@5775fe2 fix(auth): serialize OAuth refresh across agents to fix #26322 (#67876) · openclaw/openclaw@8e79080 test: allow ollama public surface boundary test · openclaw/openclaw@7d4f1a6 Docs: add test performance guardrails · openclaw/openclaw@89706d3 Tests: restore context-engine usage proof · openclaw/openclaw@e4c4f95 Tests: slim context engine runtime coverage · openclaw/openclaw@74c198f ci: retry failed custom checkouts · openclaw/openclaw@0ee5baf test: trim duplicate provider auth onboarding cases · openclaw/openclaw@1ffc02e matrix: fix sessions_spawn --thread subagent session spawning (#67643) · openclaw/openclaw@1ce2596 test: reduce auth choice fixture churn · openclaw/openclaw@857b9cd test: mock health status config boundaries · openclaw/openclaw@9d5ab4a test: mock onboard config io boundary · openclaw/openclaw@299694d test: mock legacy state plugin boundaries · openclaw/openclaw@2713089 test: mock channel install boundaries · openclaw/openclaw@b945248 test: mock doctor preview channel boundaries · openclaw/openclaw@b1a3ad4 test: trim doctor command hotspots · openclaw/openclaw@c66f16a test: isolate agent auth and spawn hotspots · openclaw/openclaw@9285935 test: stabilize MCP startup disposal race · openclaw/openclaw@dd9d2eb test: merge browser contract server suites · openclaw/openclaw@5817a76 test: narrow ollama provider discovery setup · openclaw/openclaw@a0d9598 build: declare qa-lab aimock runtime dependency · openclaw/openclaw@24431e5 test: speed up safe-bins exec harness · openclaw/openclaw@ee856ab test: preserve tool helpers in embedded runner mocks · openclaw/openclaw@acd86a0 refactor: move memory embeddings into provider plugins · openclaw/openclaw@77e6e4c test: reuse system-run temp fixtures · openclaw/openclaw@7e9ff0f test: trim hotspot wait overhead · openclaw/openclaw@12a59b0 Check: avoid duplicate boundary prep · openclaw/openclaw@baf11b8 test: reduce hotspot fixture overhead · openclaw/openclaw@3a59edd feat(ui): overhaul settings and slash command UX (#67819) thanks @Bun… · openclaw/openclaw@2cfb660 QA Matrix: exit cleanly on failure · openclaw/openclaw@42805d2 QA Matrix: isolate scenario coverage · openclaw/openclaw@7e659e1 Matrix: refresh crypto bootstrap state · openclaw/openclaw@94081d8 QA Lab: add provider registry · openclaw/openclaw@bb7e982 Matrix: add plugin changelog · openclaw/openclaw@4acab55 test: trim more hotspot overhead · openclaw/openclaw@f485311 test: trim remaining hotspot tests · openclaw/openclaw@6ba8626 test: narrow hotspot mocks · openclaw/openclaw@dbc8179 test: isolate gemini embedding request helpers · openclaw/openclaw@cd330f5 test: trim memory and mcp hotspots · openclaw/openclaw@fd48dfa test: slim provider registry mocks · openclaw/openclaw@2e08c77 test: harden Parallels update smoke · openclaw/openclaw@1a98090 feat: default Anthropic to Opus 4.7 · openclaw/openclaw@628b454 fix: harden node-host shell payload mutability checks · openclaw/openclaw@75c551e fix: land node-host approval binding for native binaries (#66731) (th… · openclaw/openclaw@29919bb CI: add daily schedule to CodeQL workflow (#67645) · openclaw/openclaw@69d25f5 fix(gateway): capture config hash after plugin auto-enable to prevent… · openclaw/openclaw@8c11210 fix: repair sanitized replay tool results before send (#67620) (thank… · openclaw/openclaw@c3c7a99 fix: restrict HTML timeout short-circuit to transient statuses · openclaw/openclaw@de129a6 fix: keep TUI watchdog bound to active run (#67401) (thanks @xantorres) · openclaw/openclaw@3525273 Gateway/skills: dedupe skills prefix-match + drop dead fallback on log · openclaw/openclaw@d7f489f Extensions/lmstudio: back off inference preload after consecutive fai… · openclaw/openclaw@b555214 TUI/streaming: add watchdog that resets the activity indicator after … · openclaw/openclaw@f44ab20 Agents/tool-loop: enable unknown-tool stream guard by default · openclaw/openclaw@36ed367 Gateway/skills: invalidate session skills snapshot on config write · openclaw/openclaw@b23d59a fix: classify HTML provider error pages correctly (#67642) (thanks @s… · openclaw/openclaw@e588e90 fix(skills): remove unused model-usage import (#67641) · openclaw/openclaw@55f05df docs(changelog): credit codex fix superseded PRs · openclaw/openclaw@e485f24 fix(openai-codex): normalize stale transport metadata in resolution a… · openclaw/openclaw@90801ba CI: pin Docker-related GitHub Actions (#67632) · openclaw/openclaw@f697b01 Android: modernize WebView and discovery API usage (#67627) · openclaw/openclaw@44a6e50 fix(deps): bump hono to 4.12.14 and @hono/node-server to 1.19.14 (GHS… · openclaw/openclaw@fbccc18 fix(deps): bump dompurify to 3.4.0 (#67614) · openclaw/openclaw@2c2dc00 CI: add explicit permissions to all workflow jobs (fixes code-scannin… · openclaw/openclaw@01b7516 fix: register bundled TTS providers and route overrides correctly (#6… · openclaw/openclaw@6ea3cdd fix: align host tilde paths with OS home (#62804) (thanks @stainlu) · openclaw/openclaw@ecfaf64 fix: flush creds queue before reconnect socket open (#67464) (thanks … · openclaw/openclaw@405c63f fix: strip standalone <function> tool call tags from visible text (#6… · openclaw/openclaw@78df859 fix(agents): preserve cli session metadata before transcript persist … · openclaw/openclaw@898fd04 docs(changelog): move cli transcript entry · openclaw/openclaw@c1817c6 fix(agents): normalize cli transcript api field · openclaw/openclaw@3a3fae0 docs(changelog): note cli transcript persistence · openclaw/openclaw@6c343f1 fix(agents): persist cli transcript turns · openclaw/openclaw@b8ef507 fix(msteams): harden security-sensitive flows (#65841) · openclaw/openclaw@c56b56e [Dashboard] Fix exec approval modal overflow for long command content… · openclaw/openclaw@053c5b0 Docs: remove QA changelog entry · openclaw/openclaw@7fd5771 QA: fix private runtime source loading (#67428) · openclaw/openclaw@d5933af docs(gateway): correct protocol.md schema path, hello-ok example, aut… · openclaw/openclaw@489404d CI: pin Node 22 runners to 22.18.0 · openclaw/openclaw@4ffa621 models.authStatus: normalize provider ids + tighten env-backed escape… · openclaw/openclaw@f2fdb9d Update CHANGELOG.md · openclaw/openclaw@7694a92 test(parallels): clean up npm update guard jobs · openclaw/openclaw@045ea7b Plugins: prefer scanDir override paths · openclaw/openclaw@b2974da fix(dreaming): default storage.mode to "separate" so phase blocks sto… · openclaw/openclaw@8c392f0 fix(memory-core): skip dreaming transcript ingestion via session stor… · openclaw/openclaw@a1b01f0 fix: dedupe replayed exec.finished node events (#67281) · openclaw/openclaw@5dcf526
fix(memory-core): keep short protected-glossary terms past the min-le… · openclaw/openclaw@bea3d29
ly-wang19 · 2026-06-24 · via Recent Commits to openclaw:main

@@ -330,7 +330,7 @@ function isKanaOnlyToken(value: string): boolean {

330330

);

331331

}

332332333-

function normalizeConceptToken(rawToken: string): string | null {

333+

function normalizeConceptToken(rawToken: string, fromGlossary = false): string | null {

334334

const normalized = normalizeLowercaseStringOrEmpty(

335335

rawToken

336336

.normalize("NFKC")

@@ -348,7 +348,9 @@ function normalizeConceptToken(rawToken: string): string | null {

348348

return null;

349349

}

350350

const script = classifyConceptTagScript(normalized);

351-

if (normalized.length < minimumTokenLengthForScript(script)) {

351+

// Glossary entries are an explicit allowlist of short technical terms (e.g. "kv", "s3"); they

352+

// bypass the per-script minimum length that would otherwise discard them.

353+

if (!fromGlossary && normalized.length < minimumTokenLengthForScript(script)) {

352354

return null;

353355

}

354356

if (isKanaOnlyToken(normalized) && normalized.length < 3) {

@@ -360,14 +362,43 @@ function normalizeConceptToken(rawToken: string): string | null {

360362

return normalized;

361363

}

362364365+

// Only entries shorter than their script's minimum token length rely on the glossary bypass, and

366+

// only those need whole-word matching so they don't fire inside longer words ("kv" in "mkv"). Longer

367+

// entries keep substring containment (the shipped behavior, e.g. "backup" tagging inside "backups").

368+

// Precomputed so derive() does not reclassify on every call.

369+

const GLOSSARY_ENTRIES = PROTECTED_GLOSSARY.map((entry) => ({

370+

entry,

371+

wholeWord: entry.length < minimumTokenLengthForScript(classifyConceptTagScript(entry)),

372+

}));

373+374+

function isAlphanumericAt(source: string, index: number): boolean {

375+

const ch = source[index];

376+

return ch !== undefined && LETTER_OR_NUMBER_RE.test(ch);

377+

}

378+379+

// True when `entry` occurs as a delimiter-bounded token, not inside a longer word. Keeps short

380+

// glossary entries like "kv"/"s3" from firing inside "mkv"/"css3" once they bypass the length gate.

381+

function includesStandaloneTerm(source: string, entry: string): boolean {

382+

let from = source.indexOf(entry);

383+

while (from !== -1) {

384+

if (!isAlphanumericAt(source, from - 1) && !isAlphanumericAt(source, from + entry.length)) {

385+

return true;

386+

}

387+

from = source.indexOf(entry, from + 1);

388+

}

389+

return false;

390+

}

391+363392

function collectGlossaryMatches(source: string): string[] {

364393

const normalizedSource = normalizeLowercaseStringOrEmpty(source.normalize("NFKC"));

365394

const matches: string[] = [];

366-

for (const entry of PROTECTED_GLOSSARY) {

367-

if (!normalizedSource.includes(entry)) {

368-

continue;

395+

for (const { entry, wholeWord } of GLOSSARY_ENTRIES) {

396+

const present = wholeWord

397+

? includesStandaloneTerm(normalizedSource, entry)

398+

: normalizedSource.includes(entry);

399+

if (present) {

400+

matches.push(entry);

369401

}

370-

matches.push(entry);

371402

}

372403

return matches;

373404

}

@@ -385,8 +416,13 @@ function collectSegmentTokens(source: string): string[] {

385416

return source.split(/[^\p{L}\p{N}]+/u).filter(Boolean);

386417

}

387418388-

function pushNormalizedTag(tags: string[], rawToken: string, limit: number): void {

389-

const normalized = normalizeConceptToken(rawToken);

419+

function pushNormalizedTag(

420+

tags: string[],

421+

rawToken: string,

422+

limit: number,

423+

fromGlossary = false,

424+

): void {

425+

const normalized = normalizeConceptToken(rawToken, fromGlossary);

390426

if (!normalized || tags.includes(normalized)) {

391427

return;

392428

}

@@ -410,14 +446,17 @@ export function deriveConceptTags(params: {

410446

}

411447412448

const tags: string[] = [];

413-

for (const rawToken of [

414-

...collectGlossaryMatches(source),

415-

...collectCompoundTokens(source),

416-

...collectSegmentTokens(source),

417-

]) {

418-

pushNormalizedTag(tags, rawToken, limit);

419-

if (tags.length >= limit) {

420-

break;

449+

const tokenSources: Array<{ tokens: string[]; fromGlossary: boolean }> = [

450+

{ tokens: collectGlossaryMatches(source), fromGlossary: true },

451+

{ tokens: collectCompoundTokens(source), fromGlossary: false },

452+

{ tokens: collectSegmentTokens(source), fromGlossary: false },

453+

];

454+

for (const { tokens, fromGlossary } of tokenSources) {

455+

for (const rawToken of tokens) {

456+

pushNormalizedTag(tags, rawToken, limit, fromGlossary);

457+

if (tags.length >= limit) {

458+

return tags;

459+

}

421460

}

422461

}

423462

return tags;