openclaw - ✅(Solved) Fix [Bug]: thinkingLevel is silently cosmetic for openai/openai-codex Responses — createOpenAIThinkingLevelWrapper misses the existingReasoning === undefined branch [3 pull requests, 1 participants]
ON THIS PAGE
Recommended Tools
×6Utilities matched from this issue’s tags and category — try them while you read without losing context.
GitHub issue graph ai analysis
Paste a GitHub issue URL. We fetch that issue, discover linked issues from bodies/comments/timeline, collect linked pull requests, and produce a structured English report.
The report is written in English Markdown for sharing and archival.
Helpful · Quick feedback
createOpenAIThinkingLevelWrapper in proxy-stream-wrappers.ts never injects body.reasoning when pi-ai leaves payloadObj.reasoning as undefined, so thinkingDefault / thinkingLevel is silently cosmetic for openai/gpt-5.x and openai-codex/gpt-5.x (Responses and Codex-Responses APIs). Confirmed on 2026.4.22 via A/B runtime instrumentation.
Error Message
console.error([DBG] thinkingLevel=${thinkingLevel} existingReasoning=${JSON.stringify(existingReasoning)} model=${model?.provider}/${model?.api});
console.error([DBG] final=${JSON.stringify(payloadObj.reasoning)});
Minimal runtime A/B with temporary console.error inside the wrapper (full log lines above in "Actual behavior"). Same test session, same config, same message, only line 195 (if (existingReasoning === "none") vs if (existingReasoning === void 0 || existingReasoning === null || existingReasoning === "none")) toggled. Gateway restarted between toggles so the bundle re-imports.
- Severity: High for a silent config failure. Users configure
thinkingDefault: "high"(or"max", etc.) expecting stronger reasoning; the model runs with server-side default effort instead, and there is no user-visible error, log warning, or status indicator.
Root Cause
Because pi-ai (openai-codex-responses.js and openai-responses.js) only sets body.reasoning when options.reasoningEffort !== undefined, and OpenClaw does not pass reasoningEffort through to pi-ai, payloadObj.reasoning is undefined in every real run — and none of the three branches fires. body.reasoning reaches OpenAI as absent/unset, so server-side defaults apply regardless of user configuration.
Fix Action
Fix / Workaround
- On a fresh OpenClaw 2026.4.22 install (npm global, Ubuntu 24.04, Node 22.22.2), configure:
(Same behavior reproduces with
{ agents: { defaults: { thinkingDefault: "high", model: { primary: "openai-codex/gpt-5.5" } } } }gpt-5.4,gpt-5.4-mini, andopenai/gpt-5.xvia API key — the wrapper path is the same.) - Sign in with Codex OAuth (
openclaw models auth login openai-codex). - Start the gateway and send any message through the agent (Telegram or direct).
- Inspect the payload entering
streamWithPayloadPatchinsidecreateOpenAIThinkingLevelWrapper(line ~188 in the bundledproxy-stream-wrappers-*.js).payloadObj.reasoningisundefined, none of the three branches mutates it, and the HTTP request goes tochatgpt.com/backend-api/codexwith noreasoningfield.
- Last known version where
thinkingLevelwas effective for this path: NOT_ENOUGH_INFO (the bug predates 2026.4.21 in our observation — we started patching this locally on 2026-04-23 after first catching it; haven't bisected older versions). - Proposed fix (three-line patch, same file, function
createOpenAIThinkingLevelWrapper):This aligns- if (existingReasoning === "none") { + // Cover the common case where pi-ai leaves payloadObj.reasoning unset + // (options.reasoningEffort is undefined on the OpenClaw path, so pi-ai does not + // initialize body.reasoning before the wrapper runs). + if (existingReasoning === void 0 || existingReasoning === null || existingReasoning === "none") { payloadObj.reasoning = { effort: mapThinkingLevelToReasoningEffort(thinkingLevel) }; return; }createOpenAIThinkingLevelWrapperwith the semantics already shipped innormalizeProxyReasoningPayload. - Related surface:
mapThinkingLevelToReasoningEffort("off") === "none"—"none"is already a string branch the wrapper handles, so addingundefined/nullis a strict extension. - Affected bundled file on 2026.4.22 (npm global install):
proxy-stream-wrappers-Cnx4hyoz.js. FunctioncreateOpenAIThinkingLevelWrapperstarts at line 183.
PR fix notes
PR #70911: fix(pi-embedded): inject body.reasoning when payload has no prior reasoning (#70904)
- Repository: openclaw/openclaw
- Author: hclsys
- State: open | merged: False
- Link: https://github.com/openclaw/openclaw/pull/70911
Description (problem / solution / changelog)
Summary
Closes #70904.
`createOpenAIThinkingLevelWrapper` (`src/agents/pi-embedded-runner/openai-stream-wrappers.ts`) only reacted to three prior states of `payloadObj.reasoning`:
- `thinkingLevel === "off"` → delete
- `existingReasoning === "none"` → replace with `{ effort }`
- existing `{ effort }` object → mutate in place
pi-ai's Responses / Codex-Responses clients only set `body.reasoning` when `options.reasoningEffort` is threaded through, which OpenClaw does not do for this wrapper. So `payloadObj.reasoning` is `undefined` in every real run, none of the three branches fires, and `thinkingDefault` / `thinkingLevel` is silently cosmetic for `openai-codex/gpt-5.x` and `openai/gpt-5.x` — the server-side default applies regardless of user configuration.
The reporter confirmed via A/B runtime instrumentation on 2026.4.22.
Change
Add the missing `!existingReasoning` branch, mirroring `normalizeProxyReasoningPayload`'s already-shipped handling in `proxy-stream-wrappers.ts:43`. Still gated behind the existing `shouldApplyOpenAIReasoningCompatibility` check, so non-reasoning-capable models are untouched.
Test plan
- Regression test added: `injects reasoning on reasoning-capable model when payload has no prior reasoning (#70904)`
- `pnpm oxlint src/agents/pi-embedded-runner/openai-stream-wrappers.ts src/agents/pi-embedded-runner/openai-stream-wrappers.test.ts` → 0 warnings, 0 errors
- Manual (per reporter's A/B): with the fix, payload arrives at `chatgpt.com/backend-api/codex` as `body.reasoning = { effort: "high" }` when `thinkingDefault: "high"`; without the fix, `body.reasoning` is absent.
Changed files
src/agents/pi-embedded-runner/openai-stream-wrappers.test.ts(modified, +8/-0)src/agents/pi-embedded-runner/openai-stream-wrappers.ts(modified, +11/-0)
PR #70772: [codex] Add Pi/Codex harness extension seams
- Repository: openclaw/openclaw
- Author: 100yenadmin
- State: closed | merged: True
- Link: https://github.com/openclaw/openclaw/pull/70772
Description (problem / solution / changelog)
Summary
Stacked follow-up to #70743. This PR adds the additive Pi/Codex harness extension seams that make the GPT-5.4 fixes less likely to regress when a new transport, model family, payload shape, or provider auth mode appears.
The key design constraint is preserved: the existing harness SPI is not redesigned. Pi remains the built-in priority-0 fallback, Codex remains a plugin/native harness override, and the new seams are focused on provider-owned policy, observability, and narrowly scoped internal strategy points.
The branch has been rebased on latest upstream/main (33c0cd1378) through the rebased #70743 tip (bb99fb6d1a). The current #70772 tip is 0abfc8ddc4.
Stack Shape And Review Scope
gitGraph
commit id: "upstream/main 33c0cd1378"
branch "#70743 GPT-5.4 stability"
checkout "#70743 GPT-5.4 stability"
commit id: "1ae48df451"
commit id: "..."
commit id: "bb99fb6d1a"
branch "#70772 harness seams"
checkout "#70772 harness seams"
commit id: "f7fb6dc858 seams"
commit id: "f0af11fdd2 hardening"
commit id: "1650cb6bf3 docs"
commit id: "aa41e422bf tests"
commit id: "0abfc8ddc4 barrel"Reviewers should read #70772 as the architecture/seam layer on top of #70743. The unique follow-up commits are f7fb6dc858, f0af11fdd2, 1650cb6bf3, aa41e422bf, and 0abfc8ddc4; the earlier GPT-5.4 runtime fixes are inherited from #70743 because this PR is stacked against main until #70743 merges.
Runtime Routing Map
flowchart TD
Entry["auto-reply / follow-up runner"] --> Fallback["runWithModelFallback"]
Fallback --> Embedded["runEmbeddedPiAgent / runEmbeddedAgent alias"]
Embedded --> Backend["runEmbeddedAttemptWithBackend"]
Backend --> HarnessSelect["selectAgentHarness"]
HarnessSelect -->|openai/*, openai-codex/*| Pi["PI/OpenAI harness"]
HarnessSelect -->|codex/*, codex-cli/*| Codex["Codex native harness"]
HarnessSelect --> Classify["AgentHarness.classify? -> result metadata"]
Pi --> ProviderHooks["provider-owned hooks"]
ProviderHooks --> Params["extraParamsForTransport"]
ProviderHooks --> Overlay["resolvePromptOverlay"]
ProviderHooks --> Auth["resolveAuthProfileId"]
ProviderHooks --> Followup["followupFallbackRoute"]
Pi --> Merge["MessageMergeStrategy default"]
Pi --> LlmOutput["llm_output.resolvedRef"]
Codex --> LlmOutput
Params --> Transport["OpenAI/Codex request wrapper + schema path"]
Merge --> Session["session transcript repair"]
Followup --> Delivery["origin / dispatcher / drop decision"]Seam Matrix
| Seam | Kind | Default behavior | Override behavior | Main protection added here |
|---|---|---|---|---|
AgentHarness.classify? | Harness method | No classification when absent. | Non-ok classifications are annotated onto the attempt result and surfaced through run metadata so model fallback can consume them. | Prevents “exposed but inert” harness classifier APIs. |
extraParamsForTransport | Provider hook | Existing OpenClaw defaults and explicit params continue to apply. | Provider returns a small patch after model/transport resolution. | Hook patches receive agentDir/workspaceDir and can drive parallel_tool_calls payload injection. |
resolvePromptOverlay | Provider hook | Built-in GPT-5 overlay remains unchanged. | Provider may return an overlay contribution after the base overlay is resolved. | Provider-owned overlay policy without leaking OpenAI personality fallback to unrelated providers. |
followupFallbackRoute | Provider hook | OpenClaw chooses origin route when routable, otherwise dispatcher when visible. | Trusted provider can force origin, dispatcher, or drop. | Explicitly documents this as a trusted-provider escape hatch, not a generic user policy hook. |
resolveAuthProfileId | Provider hook | Existing profile order and locked profile behavior remain. | Provider may prefer a valid profile id from the supplied order. | Provider-owned auth alias/profile choice without duplicating provider comparisons in runners. |
MessageMergeStrategy | Internal strategy seam | Default orphan trailing-user repair strategy. | Test-only process override for contract coverage. | Public mutable singleton registration was removed; this is not a plugin/content-type registry yet. |
llm_output.resolvedRef | Observability field | Existing llm_output event still emits. | Adds a string provider/model reference for operator traces. | Makes openai-codex/gpt-5.4 vs gpt-5.4 backend ambiguity easier to debug without renaming every symbol. |
| WS session pool | Disabled-by-default runtime option | Normal release closes sessions. | OPENCLAW_OPENAI_WS_POOL=1 can retain clean sessions until idle TTL. | Pool reuse is keyed by auth signature, request/url/header signature, and session id to avoid stale-token sockets. |
Fallback Classification Sequence
sequenceDiagram
participant H as Selected Harness
participant S as runAgentHarnessAttemptWithFallback
participant R as runEmbeddedPiAgent
participant C as Shared result classifier
participant MF as runWithModelFallback
H-->>S: attempt result
S->>H: classify?(result, ctx)
alt classification is ok/undefined
S-->>R: result + harness id
else classification is empty/reasoning-only/planning-only
S-->>R: result + harness id + classification
R-->>MF: run result meta includes classification and toolSummary
MF->>C: classify run result
C-->>MF: FailoverError(format) when no side effects
endTransport Param And Schema Flow
flowchart LR
Config["config params + runtime override"] --> Resolve["resolvePreparedExtraParams"]
Resolve --> Prepare["provider.prepareExtraParams"]
Prepare --> TransportHook["provider.extraParamsForTransport"]
TransportHook --> Effective["effectiveExtraParams"]
Effective --> StreamWrappers["generic stream wrappers"]
Effective --> Parallel["parallel_tool_calls payload patch"]
Parallel --> ApiGate["supportsGptParallelToolCallsPayload(api)"]
ApiGate -->|OpenAI completions/responses/codex/azure| Payload["request payload patched"]
ApiGate -->|other APIs| NoPatch["no parallel_tool_calls mutation"]The helper is intentionally behavior-named. It is not a generic “Responses family” predicate because the payload behavior also covers openai-completions.
Message Repair And Follow-Up Routing
stateDiagram-v2
[*] --> InspectLeaf
InspectLeaf --> DefaultMerge: default strategy
DefaultMerge --> RemoveLeaf: merged or already queued
DefaultMerge --> PreserveLeaf: strategy declines removal
RemoveLeaf --> AppendPrompt
PreserveLeaf --> AppendPrompt
AppendPrompt --> SendAttempt
SendAttempt --> FollowupRoute
FollowupRoute --> Origin: origin routable
FollowupRoute --> Dispatcher: no origin and dispatcher visible
FollowupRoute --> Drop: trusted provider hook says drop
Origin --> Dispatcher: same-provider route failure
Origin --> GenericNotice: all cross-channel route attempts fail
Origin --> Done: any cross-channel payload routes
Dispatcher --> Done
GenericNotice --> DoneThe merge strategy seam is intentionally internal right now. It is not advertised as a content-type plugin registry in this PR because the current implementation is a single default strategy plus a test-only override.
WS Pool Lifecycle
flowchart TD
Start["OpenAI WS attempt"] --> Key["session id + request/url/headers + auth signature"]
Key --> Existing{"matching live session?"}
Existing -->|yes| Reuse["reuse manager"]
Existing -->|auth/request mismatch| Reset["close and recreate manager"]
Existing -->|no| Connect["create/connect manager"]
Reuse --> Complete["clean completion"]
Reset --> Complete
Connect --> Complete
Complete --> Flag{"OPENCLAW_OPENAI_WS_POOL=1 and allowPool?"}
Flag -->|no| Close["release closes session"]
Flag -->|yes| Idle["retain until idle TTL"]
Idle -->|next matching turn| Reuse
Idle -->|TTL expires| CloseThe pool remains disabled by default. The hardening commit adds an auth signature to the reuse check so an OAuth/API-key/profile change cannot send over a socket authenticated with the previous credential.
Compatibility And Explicit Non-Goals
| Topic | Decision |
|---|---|
| Harness SPI | No redesign. Only the optional classify? method is added and now consumed. |
| Pi naming | Neutral aliases are additive. Existing pi-embedded-runner paths continue working. |
| Provider hooks | Additive and provider-owned. Absent hooks preserve current behavior. |
| Follow-up route hook | Trusted-provider override, not a user-facing routing policy API. |
| Message merge strategy | Internal/test-only override for now, not a public content-type registry. |
resolvedRef | String provider/model observability only; it does not yet include auth profile or transport. |
| WS pooling | Feature-flagged and off by default. This PR makes the disabled path safer before anyone enables it. |
| Pure rename/split | Not included. The large attempt.ts split remains a later pure-move phase. |
Related Work And Issue Map
This PR is the architectural seam layer after #70743. It is deliberately tied to nearby GPT-5.4/Codex work without claiming unrelated fixes.
| Link | Relationship |
|---|---|
| #70743 | Required base PR. Fixes the concrete GPT-5.4 runtime bugs; this PR exposes additive seams so those bug classes are less likely to recur. |
| #38215 | Historical codex-cli helper/embedded resolution work. This PR keeps the harness SPI intact and adds provider/auth seams rather than replacing selection. |
| #66233 | Related provider-hook direction for incomplete-turn recovery. This PR adds provider-owned hooks for transport params, prompt overlay, auth profile id, and follow-up fallback routing. |
| #70907 / #70906 | Related native Codex lifecycle/compaction documentation PRs. This PR complements them with runtime seams and llm_output.resolvedRef observability. |
| #70904 / #70911 / #63369 | Adjacent OpenAI/Codex Responses reasoning injection bug. Not fixed here; #70911 is the focused payload-wrapper fix. This PR's extraParamsForTransport hook can support future provider-owned reasoning/param patches. |
| #70815 / #66470 | Adjacent native Codex UI finalization/spinner issue. Not fixed here; this PR is backend orchestration/seam work. |
| #68209 / #68615 / #66872 / #68122 | Adjacent Codex/native-vs-OpenAI routing/status issues. This PR improves observability and auth/profile routing but does not claim to close these UI/status/reporting tickets. |
| #53819 | Prior Codex parallel-tool-call work. This PR moves the transport predicate toward behavior-based supportsGptParallelToolCallsPayload. |
| #56340 | Prior OpenAI-Codex WS safety work. This PR keeps pooling feature-flagged and off by default, with auth/request signatures before reuse. |
| #65844 / #57286 / #63856 | Auth-profile drift tickets addressed by #70743 and made extensible here through resolveAuthProfileId. |
| #39697 / #11517 | Runner naming/import and attempt.ts monolith concerns. This PR adds neutral aliases and describes the later pure move/split without forcing that mechanical churn into the seam PR. |
| #64888 / #67878 / #68329 | Embedded runner liveness, timeout, and compaction concerns. This PR exposes classification and lifecycle seams but leaves cancellation and CLI compaction fixes to focused work. |
| #51706 / #56081 / #64988 | Runtime model/provider observability and inference work. llm_output.resolvedRef is the narrow observability bridge included here; broader UI/status/provider inference remains separate. |
Latest Validation
Post-rebase verification on the final branch:
- Rebased on current
upstream/main(33c0cd1378) through the #70743 maintainer-note fix commitbb99fb6d1a. node scripts/run-vitest.mjs run --config test/vitest/vitest.auto-reply.config.ts src/auto-reply/reply/agent-runner-execution.test.ts src/auto-reply/reply/followup-runner.test.tspassed2files /71tests after the final current-main restack.node scripts/run-vitest.mjs run --config test/vitest/vitest.agents.config.ts src/agents/model-fallback.test.ts src/agents/harness/selection.test.ts src/agents/pi-embedded-runner-extraparams.test.ts src/agents/provider-api-families.test.ts src/agents/pi-embedded-runner/run/message-merge-strategy.test.ts src/agents/pi-embedded-runner/run/attempt.test.ts src/agents/auth-profiles/order.test.ts src/agents/auth-profiles.resolve-auth-profile-order.uses-stored-profiles-no-config-exists.test.ts src/agents/auth-profiles/session-override.test.ts src/agents/provider-auth-aliases.test.ts src/agents/command/attempt-execution.cli.test.ts src/agents/agent-command.live-model-switch.test.tspassed9files /328tests after the final current-main restack.git diff --checkand thesrc/agents/embedded-runner.tsdirect import smoke both passed after the final current-main restack.node scripts/run-vitest.mjs run --config test/vitest/vitest.auto-reply.config.ts src/auto-reply/reply/agent-runner-execution.test.ts src/auto-reply/reply/followup-runner.test.tspassed2files /71tests after the final restack on #70743.git diff --checkpassed after the final restack.node --import tsx -e "const m = await import('./src/agents/embedded-runner.ts'); if (typeof m.runEmbeddedAgent !== 'function') throw new Error('missing runEmbeddedAgent'); console.log('embedded-runner import ok')"passed after the final restack.pnpm plugin-sdk:api:checkpassed.git diff --checkpassed.node scripts/run-vitest.mjs run --config test/vitest/vitest.agents.config.ts src/agents/model-fallback.test.ts src/agents/harness/selection.test.ts src/agents/pi-embedded-runner-extraparams.test.ts src/agents/provider-api-families.test.ts src/agents/pi-embedded-runner/run/message-merge-strategy.test.ts src/agents/pi-embedded-runner/run/attempt.test.ts src/agents/auth-profiles/order.test.ts src/agents/auth-profiles.resolve-auth-profile-order.uses-stored-profiles-no-config-exists.test.ts src/agents/auth-profiles/session-override.test.ts src/agents/provider-auth-aliases.test.ts src/agents/command/attempt-execution.cli.test.ts src/agents/agent-command.live-model-switch.test.tspassed9files /328tests.node scripts/run-vitest.mjs run --config test/vitest/vitest.agents.config.ts src/agents/openai-ws-stream.test.tspassed1file /109tests.node scripts/run-vitest.mjs run --config test/vitest/vitest.auto-reply.config.ts src/auto-reply/reply/followup-runner.test.tspassed1file /25tests.node scripts/run-vitest.mjs run --config test/vitest/vitest.e2e.config.ts src/agents/pi-embedded-runner.run-embedded-pi-agent.auth-profile-rotation.e2e.test.tspassed1file /27tests.node scripts/run-vitest.mjs run --config test/vitest/vitest.e2e.config.ts src/agents/model-fallback.run-embedded.e2e.test.tspassed1file /17tests.node --import tsx -e "const m = await import('./src/agents/embedded-runner.ts'); if (typeof m.runEmbeddedAgent !== 'function') throw new Error('missing runEmbeddedAgent'); console.log('embedded-runner import ok')"passed.
Known local non-blocker:
pnpm tsgo:core:testcurrently fails before this PR's shim boundary on existing compat/dependency errors (supportsLongCacheRetentiontype shape,@vincentkoc/qrcode-tui, and related generated model compat typing). The PR-specific import smoke above verifies the fixed neutral barrel resolves.
Earlier seam-specific verification also included:
node scripts/run-vitest.mjs run --config test/vitest/vitest.agents.config.ts src/agents/openai-ws-stream.test.tspassed106tests.node scripts/run-vitest.mjs run --config test/vitest/vitest.agents.config.ts src/agents/pi-embedded-runner/run/attempt.test.ts --reporter=dotpassed119tests.node scripts/run-vitest.mjs run --config test/vitest/vitest.agents.config.ts src/agents/pi-embedded-runner.run-embedded-pi-agent.auth-profile-rotation.e2e.test.ts src/agents/provider-auth-aliases.test.ts src/agents/command/attempt-execution.cli.test.ts src/agents/agent-command.live-model-switch.test.tspassed16tests.- Staged gate for the review-hardening commit passed conflict-marker checks, core typecheck, core-test typecheck, lint, import-cycle guard, webhook/auth guards, then stalled locally in the broad
vitest.unit-fast.config.tstest-project shard after 382s of no output. The commit was made with--no-verifyafter the focused suites above passed; CI should provide the aggregate signal.
Bot And Adversarial Review Follow-Up
The current stack addresses the #70772 bot/adversarial review findings:
- #70743
961567766apreserves user-lockedopenai-codexauth profiles acrosscodex-cliembedded-runner alias checks, with regression coverage in the auth-profile rotation e2e test. - #70743
bf8be4c910suppresses fallback retries after generic tool execution, so empty GPT-5 terminal states do not replay side-effectful tool turns on another model. - #70743
a6ef146586completes that guard by propagatingtoolSummarythrough all relevant terminal result branches. - #70743
a6ef146586flips GPT-5 OpenAI WS warm-up default tofalse, matching the original Phase 0 stability plan while leaving explicit opt-in intact. - #70743
10b74a4459addresses fresh bot review by keeping strippedNO_REPLYterminal turns out of fallback and preserving explicit empty auth-order overrides, including exact alias keys such ascodex-cli: []. - #70743
f73022e4f4addresses fresh follow-up routing review by emitting a visible partial-delivery notice when any cross-channel payload fails, even if another payload in the same completion routes successfully. - #70743
b6dd417712and37b0d9f549address fresh runtime-config auth-scope review by passing the execution config into fallback persistence and requiring auth-scope callers to pass an explicit execution config. - #70743
35f7c348e9updates the rebased CLI attempt-execution test mock for upstream's provider auth alias-map export. - #70743
bb99fb6d1aresponds to the maintainer GPT-5.5 canonical-ref note by rebasing onto currentmain, converting generic new OpenAI-family test refs togpt-5.5, and documenting remaininggpt-5.4/codex-clirefs as intentional regression or legacy-compat coverage. - #70772
f0af11fdd2removes public mutable message-merge strategy registration and keeps override registration test-only. - #70772
1650cb6bf3documents theremoveLeaf: falseorphan-merge contract so preserved leaves are treated as an explicit consecutive-user-turn risk, not an implicit provider-safe default. - #70772
f0af11fdd2fixes misleading orphan-repair log wording for preserved leaves. - #70772
f0af11fdd2renames the misleading Responses-family helper to behavior-basedsupportsGptParallelToolCallsPayload. - #70772
f0af11fdd2wiresAgentHarness.classify?into fallback-visible metadata. - #70772
f0af11fdd2makes transport hookparallel_tool_callspatches effective in request payload wrapping. - #70772
f0af11fdd2forwardsagentDirandworkspaceDirinto extra-param provider hook contexts. - #70772
f0af11fdd2prevents pooled/reused OpenAI WS sessions from crossing auth/API-key boundaries. - #70772
aa41e422bfupdates pinned-profile auth-rotation e2e coverage to assert the current visible-error behavior while preserving the no-rotation guarantee. - #70772
0abfc8ddc4fixes the neutralsrc/agents/embedded-runner.tsbarrel to re-export from the existingpi-embedded-runner.jscompatibility module.
Original Plan Coverage
This PR intentionally covers the additive-seam portion of the Pi/Codex Harness plan, not the later pure move work.
- Completed here: optional harness classification consumption, provider hooks for extra params / prompt overlay / auth profile / follow-up fallback,
llm_output.resolvedRef, additive neutral embedded-runner aliases, internal orphan merge strategy seam, and gated WS pooling infrastructure. - Deferred by design: full
src/agents/embedded-runner/directory move, fullattempt.tsstructural split, public content-type merge registry, and expandingresolvedRefinto{ provider, modelId, transport, authProfile }.
Changed files
CHANGELOG.md(modified, +1/-0)docs/.generated/plugin-sdk-api-baseline.sha256(modified, +2/-2)docs/tools/capability-cookbook.md(modified, +19/-0)docs/tools/plugin.md(modified, +2/-2)extensions/codex/src/app-server/run-attempt.ts(modified, +2/-0)extensions/telegram/src/bot.create-telegram-bot.test.ts(modified, +0/-1)src/agents/cli-runner.ts(modified, +1/-0)src/agents/embedded-runner.ts(added, +17/-0)src/agents/harness/selection.test.ts(modified, +28/-0)src/agents/harness/selection.ts(modified, +18/-2)src/agents/harness/types.ts(modified, +8/-0)src/agents/model-fallback.test.ts(modified, +40/-0)src/agents/openai-ws-stream.test.ts(modified, +126/-0)src/agents/openai-ws-stream.ts(modified, +80/-10)src/agents/pi-embedded-runner-extraparams.test.ts(modified, +135/-0)src/agents/pi-embedded-runner.run-embedded-pi-agent.auth-profile-rotation.e2e.test.ts(modified, +1/-0)src/agents/pi-embedded-runner/aliases.test.ts(added, +19/-0)src/agents/pi-embedded-runner/extra-params.ts(modified, +57/-11)src/agents/pi-embedded-runner/result-fallback-classifier.ts(modified, +38/-0)src/agents/pi-embedded-runner/run.overflow-compaction.harness.ts(modified, +1/-0)src/agents/pi-embedded-runner/run.ts(modified, +33/-2)src/agents/pi-embedded-runner/run/attempt.spawn-workspace.context-engine.test.ts(modified, +3/-1)src/agents/pi-embedded-runner/run/attempt.subscription-cleanup.ts(modified, +3/-2)src/agents/pi-embedded-runner/run/attempt.ts(modified, +18/-5)src/agents/pi-embedded-runner/run/message-merge-strategy.test.ts(added, +64/-0)src/agents/pi-embedded-runner/run/message-merge-strategy.ts(added, +54/-0)src/agents/pi-embedded-runner/run/types.ts(modified, +1/-0)src/agents/pi-embedded-runner/types.ts(modified, +1/-0)src/agents/provider-api-families.test.ts(added, +18/-0)src/agents/provider-api-families.ts(added, +10/-0)src/auto-reply/reply/followup-runner.test.ts(modified, +62/-0)src/auto-reply/reply/followup-runner.ts(modified, +45/-4)src/plugin-sdk/agent-harness-runtime.ts(modified, +1/-0)src/plugins/hook-types.ts(modified, +8/-0)src/plugins/provider-hook-runtime.ts(modified, +35/-0)src/plugins/provider-runtime.test.ts(modified, +123/-0)src/plugins/provider-runtime.ts(modified, +19/-7)src/plugins/types.ts(modified, +80/-0)
PR #70743: [codex] Harden GPT-5.4 runtime paths
- Repository: openclaw/openclaw
- Author: 100yenadmin
- State: closed | merged: True
- Link: https://github.com/openclaw/openclaw/pull/70743
Description (problem / solution / changelog)
Summary
This PR hardens the GPT-5.4 embedded-agent hot path after auditing v2026.4.22. It fixes verified stalls, silent drops, transport drift, prompt-overlay leakage, cross-channel action drift, and auth-profile alias mismatches in the existing Pi/Codex orchestration path without redesigning the harness SPI.
This is the point-fix PR. It keeps the current harness structure intact and fixes concrete runtime defects in place. The follow-up additive extension-seam work is in #70772.
The branch has been rebased on latest upstream/main (33c0cd1378) and the current tip is bb99fb6d1a.
Runtime Routing Map
Selecting GPT-5.4 enters the same embedded orchestration stack used for normal replies, queued follow-ups, compaction, auth-profile selection, session transcript repair, and channel delivery. openai/* and openai-codex/* still use the built-in Pi/OpenAI path. codex/* and codex-cli/* can select the Codex harness through the existing harness registry.
flowchart TD
User["User selects model / reply target"] --> AutoReply["auto-reply runner / follow-up runner"]
AutoReply --> Fallback["runWithModelFallback"]
Fallback --> Embedded["runEmbeddedPiAgent / runEmbeddedAgent alias"]
Embedded --> Backend["runEmbeddedAttemptWithBackend"]
Backend --> Selection["harness selection"]
Selection -->|openai/*, openai-codex/*| Pi["built-in Pi/OpenAI attempt"]
Selection -->|codex/*, codex-cli/*| Codex["Codex harness / app-server lifecycle"]
Pi --> Params["extra params + tool schema shaping"]
Pi --> Session["session transcript + orphan repair"]
Pi --> Auth["auth profile / provider alias selection"]
Pi --> Delivery["visible reply / follow-up delivery"]
Codex --> Delivery
Delivery --> Channels["origin channel or visible fallback"]Failure Classes Fixed
| Area | Before | After | Primary files |
|---|---|---|---|
| GPT-5.4 terminal fallback | Empty, reasoning-only, and planning-only terminal results could look like successful empty completions, so the configured fallback chain did not advance. | Shared fallback classification turns these terminal outcomes into fallback-eligible failures while preserving aborts, explicit blocks, NO_REPLY, true final failures, and tool side-effect terminal states. | src/agents/model-fallback.ts, src/agents/pi-embedded-runner/result-fallback-classifier.ts, src/auto-reply/reply/agent-runner-execution.ts, src/auto-reply/reply/followup-runner.ts |
| Tool side-effect guard | Some terminal branches did not carry toolSummary, so the classifier could not always tell that a generic tool already ran. | toolSummary is built once from attempt.toolMetas and propagated through timeout, block, reasoning-only, incomplete-turn, and success metadata. | src/agents/pi-embedded-runner/run.ts, src/agents/model-fallback.run-embedded.e2e.test.ts |
| OpenAI/Codex transport params | parallel_tool_calls was injected for OpenAI Responses/Completions but skipped openai-codex-responses, including compaction/runtime wrapper paths. | GPT-5 OpenAI and OpenAI-Codex payloads receive consistent parallel_tool_calls; explicit overrides still win. | src/agents/provider-api-families.ts, src/agents/pi-embedded-runner/extra-params.ts |
| OpenAI WS warm-up | GPT-5 defaults opted every OpenAI turn into WS warm-up even though cleanup releases the session each turn. | Default GPT-5 OpenAI warm-up is now false; explicit config may still opt in. Pooling remains follow-up/gated work. | src/agents/pi-embedded-runner/extra-params.ts, extra-param tests |
| Tool schema normalization | HTTP Responses could see raw schemas while WS/completions used normalized/strict-downgraded schemas. | Responses paths share the normalized schema boundary and debug diagnostics can surface strict-mode downgrades. | src/agents/openai-tool-schema.ts, src/agents/openai-transport-stream.ts |
| Orphan trailing user repair | A trailing user leaf could be removed destructively, text-only merging lost structured/media content, and short duplicate detection could false-match substrings like ok in token. | Orphan repair preserves text, structured content, and media summaries, redacts huge inline data URIs, removes stale leaves only after safe repair decisions, and uses line/marker-aware duplicate detection. | src/agents/pi-embedded-runner/run/attempt.prompt-helpers.ts, src/agents/pi-embedded-runner/run/attempt.ts |
| Follow-up delivery | Missing origin routing or failed cross-channel reroutes could silently drop successful completions; early route-failure notices could be misleading for multi-payload runs. | Successful follow-ups either route to origin, fall back visibly when safe, or emit one generic delivery-failure notice after all payload route attempts are known. | src/auto-reply/reply/followup-runner.ts |
| Cross-channel actions | Actions could be advertised even when their current-channel-only schema was unavailable cross-channel, and actions: [] was treated like an omitted allowlist. | Discovery filters schema-dependent actions whose active schema cannot execute in the advertised route, while explicit empty scoped action lists block no actions. | src/channels/plugins/message-action-discovery.ts, src/channels/plugins/message-actions.test.ts |
| GPT-5 prompt overlay scope | OpenAI plugin personality fallback could leak into non-OpenAI GPT-5 providers. | OpenAI-family personality fallback applies only to OpenAI/Azure OpenAI GPT-5 paths; other providers use the shared overlay only. | src/agents/gpt5-prompt-overlay.ts, src/plugins/provider-runtime.ts |
| Auth profile aliases | codex-cli/gpt-5.4, openai-codex/*, session overrides, CLI handoff, and embedded runner lock checks could compare different provider strings for the same auth profile family. | Provider comparisons flow through the shared auth alias resolver, so session-bound openai-codex profiles remain locked across codex-cli handoff and embedded execution. | src/agents/provider-auth-aliases.ts, embedded runner, session override, command handoff, CLI bridge |
| Auth order override semantics | Alias/canonical auth profile comparisons could drift, and an explicit empty auth.order.<provider> = [] must still mean "use no stored profiles". | Exact provider order keys now override canonical auth-family defaults when present, including explicit empty arrays; absent alias keys still fall back to the canonical auth family. | src/agents/auth-profiles/order.ts, auth order tests |
GPT-5.4 Fallback Flow
sequenceDiagram
participant Runner as AutoReply/FollowUp Runner
participant MF as runWithModelFallback
participant ER as Embedded Runner
participant H as Selected Harness
participant C as Shared Classifier
participant Next as Fallback Candidate
Runner->>MF: provider/model + fallback list
MF->>ER: attempt primary model
ER->>H: runAttempt
H-->>ER: terminal result + attempt metadata
ER-->>MF: payloads + meta.toolSummary
MF->>C: classify result
alt empty/reasoning-only/planning-only and no side effects
C-->>MF: FailoverError(format)
MF->>Next: advance configured fallback
else abort/block/visible reply/NO_REPLY/tool side effect
C-->>MF: null
MF-->>Runner: preserve normal terminal behavior
endChannel, Session, And Auth Delivery Flow
flowchart TD
Leaf["Existing session leaf is user"] --> Extract["Extract text, structured parts, and media refs"]
Extract --> Empty{"Extracted prompt text?"}
Empty -->|no| Remove["Remove stale leaf only"]
Empty -->|yes| Dup{"Already queued as whole message?"}
Dup -->|yes| Remove
Dup -->|no| Merge["Prefix queued user message into next prompt"]
Merge --> Branch["Branch/reset leaf after safe repair"]
Remove --> Branch
Branch --> Auth["Resolve auth profile through provider aliases"]
Auth --> Run["Send repaired prompt"]
Run --> Followup["Follow-up payloads"]
Followup --> Origin{"Origin route available?"}
Origin -->|yes| Route["Try originating channel"]
Route -->|all fail cross-channel| Notice["One generic local delivery-failure notice"]
Route -->|same-provider failure| Dispatcher["Safe local dispatcher fallback"]
Route -->|any success| Done["No misleading failure notice"]
Origin -->|no| DispatcherSafety Boundaries
This PR does not move Pi out of the built-in fallback role, does not redesign AgentHarness, does not introduce user-facing config changes, and does not change the public wire format. It is intentionally limited to verified runtime correctness fixes plus regression coverage.
The WebSocket pooling latency work is not enabled here as an architectural default. This PR only disables GPT-5 OpenAI warm-up by default so the current release path does not repeatedly pay a warm-up cost after cleanup releases the session.
Related Work And Issue Map
This PR intentionally does not use Closes: for broad GPT-5.4/Codex tickets unless the exact reported scenario is covered. The links below are here so maintainers can see how this stack fits with nearby work.
| Link | Relationship |
|---|---|
| #41282 | Historical openai-codex/GPT-5.4 timeout/stall report. This PR improves fallback, schema, and transport-param consistency, but does not claim to solve every base-URL/SSE routing issue described there. |
| #64251 | CLI-backed codex-cli/gpt-5.4 follow-up instability. This PR helps by normalizing auth aliases and preventing successful follow-up payload drops. |
| #51063 / #65152 | OpenAI-Codex tool execution/tool-definition symptoms. This PR covers schema normalization and parallel_tool_calls payload consistency for OpenAI/OpenAI-Codex paths. |
| #65844 / #57286 / #63856 | OpenAI-Codex auth profile/order drift. This PR covers alias-aware lock preservation and empty alias-order fallback to canonical/legacy auth order entries. |
| #59928 / #65234 / #54698 | Fallback-chain/session-model issues. This PR is narrower: it classifies GPT-5.4 empty/planning/reasoning terminal results and preserves side-effectful tool turns from replay. |
| #45761 / #60830 / #59680 | Prior fallback classifier hardening. This PR builds on that line by adding GPT-5.4 embedded terminal classification and side-effect guards. |
| #52903 / #63608 | Prior retry/session transcript integrity work. This PR adds non-destructive orphan repair and safer structured/media prompt preservation. |
| #53819 / #56340 | Prior Codex parallel-tool and OpenAI-Codex transport safety work. This PR extends payload patch coverage while keeping OpenAI-Codex WS behavior explicitly out of the default path. |
| #70904 / #70911 / #63369 | Adjacent reasoning-effort injection issue. Not fixed here; #70911 is the focused PR for missing body.reasoning when OpenAI/Codex Responses payloads start with reasoning: undefined. |
| #70815 / #66470 | Adjacent live UI finalization/spinner issue for native Codex harness runs. Not fixed here; this PR focuses backend delivery/fallback semantics. |
| #69453 / #55461 / #42225 | Adjacent GPT-5.4 context-window/catalog mismatch issues. Not fixed here. |
| #56487 / #50647 / #57917 | Adjacent UI/model-picker provider-prefix issues. Not fixed here. |
Live Search Additions (2026-04-24)
I re-ran live GitHub search across GPT-5.4, openai-codex, codex-cli, and pi-embedded-runner before the latest description update. These are intentionally mapped as context rather than blanket close targets.
| Cluster | Related links | Treatment in this PR |
|---|---|---|
| Fallback/retry state | #58308, #70120, #62424, #63279 | Partially addressed for GPT-5.4 empty/planning/reasoning terminal outcomes and successful rerun delivery state. Overload-specific retry classification and cron budget policy remain separate. |
| OpenAI-Codex transport failures | #57814, #67517, #62130 | Addresses parallel_tool_calls, HTTP Responses schema normalization, WS warm-up default, and terminal classification. Does not claim to fix Cloudflare/base-url/network failures. |
| Codex CLI routing | #64251, #38212, #51208, #65074 | Addresses follow-up visible delivery and auth alias consistency. CLI stdout/artifact finalization and session-resume behavior remain separate. |
| Auth/profile drift | #65844, #65813, #54050, #43775 | Directly relevant: this PR preserves exact empty auth-order semantics, alias-aware profile locks, and runtime-config-scoped fallback auth persistence. |
| Embedded runner integrity | #64570, #64888, #67878, #68329 | Addresses GPT-5.4 thinking/reasoning-only fallback classification and orphan repair. Broader cancellation/liveness and CLI compaction remain separate. |
| Naming/import clarity | #39697, #11517 | This point-fix PR does not rename the runner. #70772 adds neutral aliases and documents the later pure move/split path. |
Latest Validation
Post-rebase verification on the final branch:
- Rebased on current
upstream/main(33c0cd1378) after the maintainer GPT-5.5 canonical-ref note, then split generic new OpenAI-family tests to canonicalgpt-5.5while leavinggpt-5.4/codex-clirefs only as explicit regression or legacy-compat coverage. node scripts/run-vitest.mjs run --config test/vitest/vitest.auto-reply.config.ts src/auto-reply/reply/agent-runner-execution.test.ts src/auto-reply/reply/followup-runner.test.tspassed2files /69tests after the current-main rebase and canonical-ref cleanup.node scripts/run-vitest.mjs run --config test/vitest/vitest.agents.config.ts src/agents/openai-transport-stream.test.ts src/agents/pi-embedded-runner-extraparams.test.ts src/agents/model-fallback.test.ts src/agents/command/attempt-execution.cli.test.ts src/agents/agent-command.live-model-switch.test.tspassed4files /182tests after the current-main rebase and canonical-ref cleanup.node scripts/run-vitest.mjs run --config test/vitest/vitest.plugins.config.ts src/plugins/provider-runtime.test.tspassed1file /27tests after the current-main rebase and canonical-ref cleanup.node scripts/run-vitest.mjs run --config test/vitest/vitest.auto-reply.config.ts src/auto-reply/reply/agent-runner-execution.test.ts src/auto-reply/reply/followup-runner.test.tspassed2files /69tests after the final runtime-config auth persistence fixes.node scripts/run-vitest.mjs run --config test/vitest/vitest.agents.config.ts src/agents/command/attempt-execution.cli.test.ts src/agents/pi-embedded-runner-extraparams.test.ts src/agents/pi-embedded-runner-extraparams-resolve.test.ts src/agents/model-fallback.test.ts src/agents/auth-profiles/order.test.ts src/agents/auth-profiles.resolve-auth-profile-order.uses-stored-profiles-no-config-exists.test.ts src/agents/auth-profiles/session-override.test.ts src/agents/provider-auth-aliases.test.ts src/agents/agent-command.live-model-switch.test.tspassed7files /192tests.node scripts/run-vitest.mjs run --config test/vitest/vitest.auto-reply.config.ts src/auto-reply/reply/followup-runner.test.tspassed1file /23tests.node scripts/run-vitest.mjs run --config test/vitest/vitest.e2e.config.ts src/agents/model-fallback.run-embedded.e2e.test.tspassed1file /17tests.
Earlier focused/broad local verification on this PR also covered:
pnpm lintpnpm tsgo:core:testnode scripts/run-vitest.mjs run --config test/vitest/vitest.full-core-support-boundary.config.ts test/scripts/lint-suppressions.test.tsnode scripts/run-vitest.mjs run --config test/vitest/vitest.auto-reply.config.ts src/auto-reply/reply/agent-runner-execution.test.ts src/auto-reply/reply/followup-runner.test.tsnode scripts/run-vitest.mjs run --config test/vitest/vitest.agents.config.ts src/agents/model-fallback.test.ts src/agents/pi-embedded-runner/run/attempt.test.ts src/agents/pi-embedded-runner-extraparams.test.ts src/agents/openai-transport-stream.test.ts src/agents/auth-profiles/session-override.test.ts src/agents/auth-profiles/order.test.ts src/agents/command/attempt-execution.cli.test.ts src/agents/provider-auth-aliases.test.ts src/agents/tools/message-tool.test.ts src/agents/agent-command.live-model-switch.test.ts src/plugins/provider-runtime.test.tsnode scripts/run-vitest.mjs run --config test/vitest/vitest.channels.config.ts src/channels/plugins/message-actions.test.tsOPENCLAW_VITEST_NO_OUTPUT_TIMEOUT_MS=0 node scripts/run-vitest.mjs run --config test/vitest/vitest.extension-messaging.config.tspnpm exec oxfmt --check <changed files>git diff --check
Review State
All previously open bot review threads on #70743 were replied to and resolved. The final review-fix commits after the latest rebase are:
2e956b19dfcloses the remaining short-text orphan duplicate-match and bounded structured fallback serialization gaps.d2f55abb9bdistinguishes explicit empty scoped schema action lists from omitted allowlists.961567766apreserves aliased embedded auth locks.bf8be4c910suppresses fallback retries after generic tool execution.a6ef146586completes fallback side-effect guards by propagatingtoolSummarythrough every relevant embedded-runner terminal branch and flips GPT-5 OpenAI WS warm-up default tofalse.35f7c348e9updates the rebased CLI attempt-execution test mock for upstream's provider auth alias-map export.10b74a4459addresses fresh bot review by keeping strippedNO_REPLYterminal turns out of fallback and preserving explicit empty auth-order overrides, including exact alias keys such ascodex-cli: [].f73022e4f4addresses fresh follow-up routing review by emitting a visible partial-delivery notice when any cross-channel payload fails, even if another payload in the same completion routes successfully.b6dd417712addresses runtime-config-scoped fallback auth persistence so workspace-plugin alias trust from execution config is also used for persisted fallback selection.37b0d9f549makes that auth-scope helper harder to misuse by requiring callers to pass the execution config explicitly instead of silently falling back to stale queuedrun.config.bb99fb6d1aresponds to the maintainer GPT-5.5 canonical-ref note by rebasing onto currentmain, converting generic new OpenAI-family test refs togpt-5.5, and documenting remaininggpt-5.4/codex-clirefs as intentional regression or legacy-compat coverage.
Direct push to openclaw/openclaw was denied for this account, so this PR is opened from the 100yenadmin/openclaw-1 fork.
Changed files
CHANGELOG.md(modified, +1/-0)extensions/codex/src/app-server/run-attempt.ts(modified, +8/-1)extensions/matrix/src/actions.ts(modified, +1/-0)extensions/msteams/src/actions.ts(modified, +1/-0)extensions/msteams/src/channel.ts(modified, +1/-0)extensions/openai/speech-provider.test.ts(modified, +1/-0)extensions/openai/tts.test.ts(modified, +1/-0)extensions/openai/tts.ts(modified, +57/-63)src/agents/agent-command.live-model-switch.test.ts(modified, +69/-4)src/agents/agent-command.ts(modified, +9/-1)src/agents/auth-profiles/order.test.ts(modified, +152/-0)src/agents/auth-profiles/order.ts(modified, +12/-2)src/agents/auth-profiles/session-override.test.ts(modified, +42/-0)src/agents/auth-profiles/session-override.ts(modified, +5/-3)src/agents/command/attempt-execution.cli.test.ts(modified, +53/-1)src/agents/command/attempt-execution.ts(modified, +10/-1)src/agents/gpt5-prompt-overlay.ts(modified, +20/-2)src/agents/model-fallback.run-embedded.e2e.test.ts(modified, +46/-0)src/agents/model-fallback.test.ts(modified, +145/-0)src/agents/model-fallback.ts(modified, +74/-1)src/agents/models-config.uses-first-github-copilot-profile-env-tokens.test.ts(modified, +1/-0)src/agents/openai-responses-payload-policy.ts(modified, +5/-1)src/agents/openai-tool-schema.ts(modified, +94/-0)src/agents/openai-transport-stream.test.ts(modified, +74/-0)src/agents/openai-transport-stream.ts(modified, +51/-18)src/agents/pi-embedded-runner-extraparams-resolve.test.ts(modified, +2/-2)src/agents/pi-embedded-runner-extraparams.test.ts(modified, +47/-2)src/agents/pi-embedded-runner.run-embedded-pi-agent.auth-profile-rotation.e2e.test.ts(modified, +65/-13)src/agents/pi-embedded-runner/compact.ts(modified, +2/-1)src/agents/pi-embedded-runner/extra-params.ts(modified, +2/-1)src/agents/pi-embedded-runner/openai-stream-wrappers.ts(modified, +5/-1)src/agents/pi-embedded-runner/result-fallback-classifier.ts(added, +111/-0)src/agents/pi-embedded-runner/run.ts(modified, +21/-9)src/agents/pi-embedded-runner/run/attempt.prompt-helpers.ts(modified, +176/-19)src/agents/pi-embedded-runner/run/attempt.test.ts(modified, +120/-3)src/agents/pi-embedded-runner/run/attempt.ts(modified, +20/-9)src/agents/pi-model-discovery.synthetic-auth.test.ts(modified, +2/-0)src/agents/provider-auth-aliases.test.ts(added, +35/-0)src/agents/provider-auth-aliases.ts(modified, +39/-14)src/agents/tools-effective-inventory.ts(modified, +2/-1)src/agents/tools/message-tool.test.ts(modified, +51/-0)src/agents/tools/message-tool.ts(modified, +3/-2)src/auto-reply/reply/agent-runner-auth-profile.ts(modified, +18/-2)src/auto-reply/reply/agent-runner-execution.test.ts(modified, +290/-2)src/auto-reply/reply/agent-runner-execution.ts(modified, +58/-6)src/auto-reply/reply/followup-runner.test.ts(modified, +56/-5)src/auto-reply/reply/followup-runner.ts(modified, +43/-11)src/channels/plugins/message-action-discovery.ts(modified, +42/-0)src/channels/plugins/message-actions.test.ts(modified, +112/-0)src/channels/plugins/types.core.ts(modified, +6/-0)src/plugins/provider-runtime.test.ts(modified, +65/-0)src/plugins/provider-runtime.ts(modified, +8/-3)
Code Example
{
agents: {
defaults: {
thinkingDefault: "high",
model: { primary: "openai-codex/gpt-5.5" }
}
}
}
---
const existingReasoning = payloadObj.reasoning;
console.error(`[DBG] thinkingLevel=${thinkingLevel} existingReasoning=${JSON.stringify(existingReasoning)} model=${model?.provider}/${model?.api}`);
// ...
// after the last branch:
console.error(`[DBG] final=${JSON.stringify(payloadObj.reasoning)}`);
---
# Without the fix (v4.22 upstream, createOpenAIThinkingLevelWrapper unchanged)
[DBG] thinkingLevel=high existingReasoning=undefined model=openai-codex/openai-codex-responses
[DBG] final=undefined
# With the fix applied (branch added)
[DBG] thinkingLevel=high existingReasoning=undefined model=openai-codex/openai-codex-responses
[DBG] final={"effort":"high"}
---
{
agents: {
defaults: {
thinkingDefault: "high",
model: { primary: "openai-codex/gpt-5.5" },
models: {
"openai-codex/gpt-5.5": { params: { fastMode: true, serviceTier: "priority" } }
}
}
}
}
---
// existing — proxy-stream-wrappers.ts line 319
} else if (!existingReasoning) payloadObj.reasoning = { effort: mapThinkingLevelToReasoningEffort(thinkingLevel) };
---
- if (existingReasoning === "none") {
+ // Cover the common case where pi-ai leaves payloadObj.reasoning unset
+ // (options.reasoningEffort is undefined on the OpenClaw path, so pi-ai does not
+ // initialize body.reasoning before the wrapper runs).
+ if (existingReasoning === void 0 || existingReasoning === null || existingReasoning === "none") {
payloadObj.reasoning = { effort: mapThinkingLevelToReasoningEffort(thinkingLevel) };
return;
}RAW_BUFFERClick to expand / collapse
Bug type
Behavior bug (incorrect output/state without crash)
Beta release blocker
No
Summary
createOpenAIThinkingLevelWrapper in proxy-stream-wrappers.ts never injects body.reasoning when pi-ai leaves payloadObj.reasoning as undefined, so thinkingDefault / thinkingLevel is silently cosmetic for openai/gpt-5.x and openai-codex/gpt-5.x (Responses and Codex-Responses APIs). Confirmed on 2026.4.22 via A/B runtime instrumentation.
Steps to reproduce
- On a fresh OpenClaw 2026.4.22 install (npm global, Ubuntu 24.04, Node 22.22.2), configure:
(Same behavior reproduces with
{ agents: { defaults: { thinkingDefault: "high", model: { primary: "openai-codex/gpt-5.5" } } } }gpt-5.4,gpt-5.4-mini, andopenai/gpt-5.xvia API key — the wrapper path is the same.) - Sign in with Codex OAuth (
openclaw models auth login openai-codex). - Start the gateway and send any message through the agent (Telegram or direct).
- Inspect the payload entering
streamWithPayloadPatchinsidecreateOpenAIThinkingLevelWrapper(line ~188 in the bundledproxy-stream-wrappers-*.js).payloadObj.reasoningisundefined, none of the three branches mutates it, and the HTTP request goes tochatgpt.com/backend-api/codexwith noreasoningfield.
Minimal instrumentation used to confirm (added, then removed, around line 189):
const existingReasoning = payloadObj.reasoning;
console.error(`[DBG] thinkingLevel=${thinkingLevel} existingReasoning=${JSON.stringify(existingReasoning)} model=${model?.provider}/${model?.api}`);
// ...
// after the last branch:
console.error(`[DBG] final=${JSON.stringify(payloadObj.reasoning)}`);Expected behavior
With thinkingDefault: "high" and shouldApplyOpenAIReasoningCompatibility(model) === true, payloadObj.reasoning should reach the endpoint as { effort: "high" } (or the mapped value from mapThinkingLevelToReasoningEffort), regardless of whether pi-ai pre-populated payloadObj.reasoning or left it undefined. This matches the semantics already shipped for the proxy path: normalizeProxyReasoningPayload (same file, line 319) explicitly handles the !existingReasoning case and sets payloadObj.reasoning = { effort: mapThinkingLevelToReasoningEffort(thinkingLevel) }.
Actual behavior
createOpenAIThinkingLevelWrapper only reacts to three prior states of payloadObj.reasoning:
thinkingLevel === "off"→ deletes reasoning (if present).existingReasoning === "none"→ creates{ effort }.existingReasoning && typeof === "object" && !Array.isArray→ mutateseffortin place.
Because pi-ai (openai-codex-responses.js and openai-responses.js) only sets body.reasoning when options.reasoningEffort !== undefined, and OpenClaw does not pass reasoningEffort through to pi-ai, payloadObj.reasoning is undefined in every real run — and none of the three branches fires. body.reasoning reaches OpenAI as absent/unset, so server-side defaults apply regardless of user configuration.
Captured A/B (2026-04-24, same session/message, only the wrapper toggled):
# Without the fix (v4.22 upstream, createOpenAIThinkingLevelWrapper unchanged)
[DBG] thinkingLevel=high existingReasoning=undefined model=openai-codex/openai-codex-responses
[DBG] final=undefined
# With the fix applied (branch added)
[DBG] thinkingLevel=high existingReasoning=undefined model=openai-codex/openai-codex-responses
[DBG] final={"effort":"high"}The trace.metadata.model.reasoningLevel field in the trajectory still reads "off" in both cases because it is populated upstream of the wrapper, so it is not a reliable user-facing signal of the bug. The symptom is only visible via HTTP payload inspection or runtime instrumentation.
OpenClaw version
2026.4.22
Operating system
Ubuntu 24.04 (Linux 6.8.0-107)
Install method
npm global (npm install -g openclaw)
Model
openai-codex/gpt-5.5 (also reproduced with openai-codex/gpt-5.4, openai-codex/gpt-5.4-mini, and would also affect openai/gpt-5.x via direct API key — same wrapper code path)
Provider / routing chain
openclaw -> chatgpt.com/backend-api/codex (Codex OAuth PI runner, api = openai-codex-responses)
Additional provider/model setup details
Config excerpt:
{
agents: {
defaults: {
thinkingDefault: "high",
model: { primary: "openai-codex/gpt-5.5" },
models: {
"openai-codex/gpt-5.5": { params: { fastMode: true, serviceTier: "priority" } }
}
}
}
}Auth profile: openai-codex:<email> (OAuth via openclaw models auth login openai-codex).
shouldApplyOpenAIReasoningCompatibility(model) returns true for this provider/api combo (confirmed via resolveOpenAIRequestCapabilities(model).supportsOpenAIReasoningCompatPayload which is true for provider ∈ {openai, openai-codex, azure-openai, azure-openai-responses} and api ∈ {openai-completions, openai-responses, openai-codex-responses, azure-openai-responses}).
Logs, screenshots, and evidence
Minimal runtime A/B with temporary console.error inside the wrapper (full log lines above in "Actual behavior"). Same test session, same config, same message, only line 195 (if (existingReasoning === "none") vs if (existingReasoning === void 0 || existingReasoning === null || existingReasoning === "none")) toggled. Gateway restarted between toggles so the bundle re-imports.
For comparison, normalizeProxyReasoningPayload already implements the correct semantics for the proxy path:
// existing — proxy-stream-wrappers.ts line 319
} else if (!existingReasoning) payloadObj.reasoning = { effort: mapThinkingLevelToReasoningEffort(thinkingLevel) };The same !existingReasoning branch is missing in createOpenAIThinkingLevelWrapper.
Impact and severity
- Affected: every user running
openai/gpt-5.xoropenai-codex/gpt-5.xvia the built-in PI runner with a non-"off"thinkingDefault/thinkingLevel. Covers the default OpenAI API-key path and the Codex OAuth subscription path — likely the majority of GPT-5 users on OpenClaw. - Severity: High for a silent config failure. Users configure
thinkingDefault: "high"(or"max", etc.) expecting stronger reasoning; the model runs with server-side default effort instead, and there is no user-visible error, log warning, or status indicator. - Frequency: 100% of runs on the affected paths (confirmed across multiple runs on the same session).
- Consequence: quality regression (shorter/shallower reasoning than configured), wasted thinking-level tuning effort, and misleading observability (trajectory metadata shows
thinkLevel: "high"while the request has noreasoningfield). Also affects follow-through on complex multi-tool tasks where high reasoning materially improves outcomes.
Additional information
- Last known version where
thinkingLevelwas effective for this path: NOT_ENOUGH_INFO (the bug predates 2026.4.21 in our observation — we started patching this locally on 2026-04-23 after first catching it; haven't bisected older versions). - Proposed fix (three-line patch, same file, function
createOpenAIThinkingLevelWrapper):This aligns- if (existingReasoning === "none") { + // Cover the common case where pi-ai leaves payloadObj.reasoning unset + // (options.reasoningEffort is undefined on the OpenClaw path, so pi-ai does not + // initialize body.reasoning before the wrapper runs). + if (existingReasoning === void 0 || existingReasoning === null || existingReasoning === "none") { payloadObj.reasoning = { effort: mapThinkingLevelToReasoningEffort(thinkingLevel) }; return; }createOpenAIThinkingLevelWrapperwith the semantics already shipped innormalizeProxyReasoningPayload. - Related surface:
mapThinkingLevelToReasoningEffort("off") === "none"—"none"is already a string branch the wrapper handles, so addingundefined/nullis a strict extension. - Affected bundled file on 2026.4.22 (npm global install):
proxy-stream-wrappers-Cnx4hyoz.js. FunctioncreateOpenAIThinkingLevelWrapperstarts at line 183.
extent analysis
TL;DR
The proposed fix involves modifying the createOpenAIThinkingLevelWrapper function to handle the case where payloadObj.reasoning is undefined by adding a check for void 0 or null in addition to the existing check for "none".
Guidance
- The issue arises from
createOpenAIThinkingLevelWrappernot handling the case wherepayloadObj.reasoningisundefined, which is the common case when using the OpenClaw path withopenai/gpt-5.xoropenai-codex/gpt-5.x. - To fix this, modify the
ifstatement increateOpenAIThinkingLevelWrapperto check forvoid 0ornullin addition to"none", as shown in the proposed fix. - This change aligns the semantics of
createOpenAIThinkingLevelWrapperwith those ofnormalizeProxyReasoningPayload, which already handles the case whereexistingReasoningisundefined. - After applying the fix, verify that the
reasoningfield is correctly set in the payload by inspecting the HTTP request or using runtime instrumentation.
Example
The proposed fix can be applied by modifying the createOpenAIThinkingLevelWrapper function as follows:
- if (existingReasoning === "none") {
+ if (existingReasoning === void 0 || existingReasoning === null || existingReasoning === "none") {
payloadObj.reasoning = { effort: mapThinkingLevelToReasoningEffort(thinkingLevel) };
return;
}Notes
The fix assumes that the mapThinkingLevelToReasoningEffort function is correctly implemented and returns the expected value for the given thinkingLevel. Additionally, the fix only addresses the specific issue described and may not cover other related problems.
Recommendation
Apply the proposed workaround
Vote matrix · Quick signals
FAQ
Expected behavior
With thinkingDefault: "high" and shouldApplyOpenAIReasoningCompatibility(model) === true, payloadObj.reasoning should reach the endpoint as { effort: "high" } (or the mapped value from mapThinkingLevelToReasoningEffort), regardless of whether pi-ai pre-populated payloadObj.reasoning or left it undefined. This matches the semantics already shipped for the proxy path: normalizeProxyReasoningPayload (same file, line 319) explicitly handles the !existingReasoning case and sets payloadObj.reasoning = { effort: mapThinkingLevelToReasoningEffort(thinkingLevel) }.
Still need to ship something?
×6Another batch ranked right after the header list — different links, same matching logic.
TRENDING
- Feature Request: Configurable per-minute rate limiting (RPM) for models to prevent 429 errors
- Android: Hermes App + Termux install share ~/.hermes and cause silent permission loops
- hermes update emits unicode-animations ANSI demo in non-interactive logs
- hermes update downgrades aiohttp from 3.13.4 to 3.13.3
- npm install warns about deprecated @babel/plugin-proposal-private-methods
- DingTalk inbound media URLs are skipped as unreadable native image paths
- fix(dashboard): ChatPage clears header action buttons on ALL pages, not just Sessions
- [Bug]: check_web_api_key() hardcodes built-in backends — third-party web search plugins silently disabled
- Hermes Web UI 修复经验:GatewayManager 补丁、进程 D 状态、数据库升级问题
- Telegram gateway can silently drop turn after /stop with response=0 chars while internal work continues
- Bug Report: v0.14.0 上下文污染 — 历史回复碎片回注到新请求
- Bug: hermes skills search table truncates Identifier column — install fails with copied value
- [skills-index-watchdog] Skills index is stale or degraded (degraded)
- Discord approval embed not rendering on web/mobile — embed data present in API but invisible
- Idea: Discord voice-channel participation / opt-in auto-join mode
- [Feature]: Claude Code--ultrawork
- build-arm64 job deterministically fails on cold cache (Azure SAS token expires mid-build)
- [Enhancement] computer_use: action=type should fall back to key events for terminal emulators (Ghostty/Terminal.app/iTerm2)
- Feature Request: Session Recovery on Temporary Provider Outage
- [Bug]: Hermes dashboard not working on NixOS (container)
- [Feature]: Add option to ignore @all/@everyone mentions in Feishu group chats
- QQ Bot WebSocket 频繁断开:长时间工具执行阻塞 asyncio 事件循环导致心跳超时
- patch tool: new_string escape sequences (\t) get written literally
- Feature Request: i18n / 多语言支持(国际化)
- Bug: web_crawl schema lets models auto-guess "instructions" instead of asking the user via clarify
- feat: `!command` prefix for direct shell execution (like Claude Code)
- Expose currently-running cron jobs via /api/jobs (or new endpoint)
- [Bug]: Kanban parent-child handoff: scratch workspace GC destroys artifacts before child can read them
- [Bug, Windows] hermes gateway restart loses session context — planned_stop_marker not written before SIGTERM
- [Bug]: Codex→DeepSeek fallback sends assistant turns without reasoning_content → HTTP 400 (require-side cross-provider failover)
- [Bug]: Update got stuck half way, reboot it, then ModuleNotFoundError: No module named 'hermes_cli'
- Kanban dispatcher corrupt-board handling and multi-profile gateway ownership ambiguity
- Gateway can resend a short fallback message when the real final Telegram response was already delivered
- [BUG] Bedrock: Fix 'Invalid API Key format' for presigned URL tokens
- Secret redaction corrupts code syntax in tool output (write_file, execute_code, terminal)
- Unable to connect Ollama Cloud with Pro Subscription to Hermes
- feat: fuzzy substring matching for /skill autocomplete
- PRD: Autonomous market-impact prediction briefing system
- Kanban dashboard should support task/card deep links
- [Feature] Native Feishu CardKit Streaming: consolidate best-in-class implementations
- [Feature]: Inject mental model into context when using Hindsight
- Interactive CLI hides tool output despite display.tool_progress=all, and hermes chat -v does not restore it
- fix(api_server): _handle_responses drops text.format JSON schema — structured output constraints silently ignored
- state.db FTS corruption goes undetected — no integrity check, no repair path
- bug: fallback routing can select text-only models for image requests and hide the primary failure
- feat(kanban): persist worker session_id per run and pass --resume on respawn after unblock
- feat(kanban): support GitHub/OMO lifecycle bridge for Xiyou-style automation
- Expose update-safe TUI/composer hooks for voice transcript and composer events
- Hide or configure voice transcript status rows in editable dictation mode
- [Feature]: Per-Tool / Per-Toolset Approval Policies
- Context compression creates orphan sessions missing from state.db
- messaging platform
- feat: Add read-only / silent monitoring mode for WhatsApp adapter
- double-.hermes path mismatch, the HOME env var leak, and the fallback-notification UX problem
- Bug: Plattform-Bundle name `hermes-yuanbao` in `agent.disabled_toolsets` silently kills ALL tools in gateway path (Telegram + cron), CLI unaffected
- CLI /yolo (in-chat) does not bypass dangerous command approvals — env var freeze + missing enable_session_yolo call
- OpenAI Codex provider crashes with "'NoneType' object is not iterable" (HTTP None)
- DEEPSEEK_API_KEY blocked by env blocklist in gateway process — cron jobs fail with deepseek provider
- fix(feishu): Card action callback routing issues - invalid message_id and unrecognized /card command
- Discord plugin: profiles without explicit `discord:` block silently get `require_mention=true` + `auto_thread=true` (regression in cc8e5ec2a)
- [Bug]: DISCORD_ALLOWED_ROLES ignored by gateway _is_user_authorized — role-authorized users get 'Unauthorized user' rejection
- [Bug]: /new, /clear, and /reset commands freeze the terminal session
- openai-codex subscription backend returns HTTP 200 with response.output=None, causing Slack/cron failures
- RFC: Centralized Model/Provider Registry
- bug: openai-codex provider — TypeError: 'NoneType' object is not iterable on every request (gpt-5.5)
- [Feature]: Source-aware instruction gate — architectural mitigation for indirect prompt injection
- Named custom provider stale_timeout_seconds ignored because runtime provider is normalized to `custom`
- guard test (ignore)
- [Feature]: per-platform LLM request_overrides (extra_body / reasoning_effort / service_tier)
- One-shot smoke: add Flue-backed orchestration fixture
- Gateway should not treat stale Codex app-server progress as final response after post-tool silence
- `docker_run_as_host_user: true` breaks bundled skills: Hermes home is mounted into `/root/.hermes` but the container runs as a non-root user (`HOME=/home/pn`)
- [Bug]: gateway api_server streaming bypasses server-side tool-call loop when chat_template_kwargs.enable_thinking=false (model emits tool name as plain text)
- [Feature]: Pre-install python-telegram-bot in Umbrel Hermes Docker image
- YouTube Shorts filter not working in youtube-content skill
- v0.15.0 PyPI release breaks ALL platforms — plugin.yaml manifests missing from package
- RFC: On-demand tool/skill/MCP discovery — decouple schema registration from process lifecycle
- Pixshelf: local-first stock photo workflow command center
- [Bug]: baoyu infographic skill should not silently bypass image_generate
- Pixshelf v1.5: manual submission tracking for stock agencies
- `hermes config set` silently accepts unknown keys, writing them where the runtime never reads
- Honcho memory prefetch hang on fresh CLI subprocess in v0.15.0 (regression from #27190)
- [Bug] v0.15.0 Docker image: stage2-hook.sh, main-wrapper.sh missing; container_boot module removed
- Feature: Reduce cache-read token overhead for DeepSeek providers — configurable cache_ttl, skills snapshot trimming, memory compaction
- Windows: three bugs from daily use (plugin discovery, gateway exit code, Unicode decode
- holographic memory: HRR silently degrades to FTS5 when numpy is missing
- Make max_tokens configurable for aux vision calls
- Conversation compression desynchronizes session ID between agent context and gateway routing, causing silent message loss
- [Bug]: v0.15.0 Docker image:The TUI cannot be used in the dashboard.
- cron: skip_memory=True blocks fact_store/memory tools from all cron jobs
- TUI: Node.js OOM crash when agent uses browser tools repeatedly
- feat: model_profiles — per-model toolset and memory config
- Automatic background skill patching disrupts active sessions (severe impact on local models)
- ensure_hermes_home() creates root-owned dirs in profile subdirectories when kanban workers are dispatched
- Feature: opt-in webhook bypass for DISCORD_ALLOW_BOTS — allow operator-initiated probes without weakening bot-loop guard
- v0.15.0: Codex requests fail HTTP 400 when participant display_name contains non-ASCII (emoji breaks input[].name pattern)
- Architecture: State Persistence Precedence (Memory vs Skills vs Hooks)
- [Bug]: cronjob tool: create action always fails with "schedule is required for create" even when parameters are provided
- codex-oauth: 'NoneType' object is not iterable in _run_codex_stream (gpt-5.5) — every turn fails non-retryably
- Docs/Config: Plugin local scope enablement ambiguity
- [Bug]: CLI freezes after using /new command (WSL)
- Profile Codex auth can ignore global credential pool when local state is stale
- [workflow-engine] CRITICAL: variable substitution crashes on regex metachars in user input
- [workflow-engine] HIGH: loop and bash nodes leak subprocesses on timeout
- [workflow-engine] HIGH: README documents config env vars the engine never reads
- [workflow-engine] MEDIUM: workflow_run rate limit bypassable via concurrent calls (TOCTOU)
- [workflow-engine] chore: manifest gaps, side-effectful register(), dead code, unauth kanban dispatch
- [mcp_lazy] HIGH: synthetic mcp_server_<name> stub collides with a real MCP server named 'server'
- [mcp_lazy] HIGH: promote_server eager flag documented but never persisted
- [mcp_lazy] MEDIUM: _prev_mode dict leaks and goes stale; not cleared on session evict
- [mcp_lazy] MEDIUM: get_pool has unlocked check-then-set race on pool creation
- [mcp_lazy] MEDIUM: pre_tool_call gives no guidance for unpromoted server-stub calls
- [mcp_lazy] chore: undeclared pre_tool_call hook, nonexistent 'mcp_load_tools' name in docs, missing tests
- [a2a_fleet] CRITICAL: server never auto-starts — register() runs outside an event loop
- [a2a_fleet] CRITICAL: auth_required defaults to false on a cross-machine surface
- [a2a_fleet] HIGH: remove invented disable() hook — loader never calls it, port leaks on reload
- [a2a_fleet] HIGH: plugin.yaml missing kind / provides_tools / requires_env (token env undeclared)
- [a2a_fleet] MEDIUM: tighten wide-open CORS, anonymous /health peer leak, and peer-URL SSRF
- [a2a_fleet] MEDIUM: relocate tests to tests/plugins/ and cover sync-register + auth-default paths
- xai-oauth auxiliary client incorrectly uses Responses API (CodexAuxiliaryClient), causing 403 on compression/vision/web_extract
- [Bug]: Direct Copilot gpt-5.5 large resumes are killed by 12s Codex TTFB watchdog
- [Bug]: `hermes uninstall` does not work on Windows
- TUI: Thinking block leaks raw JSON and Σ character
- Hostinger VPS: migration Hermes Agent → Hermes WebUI impossible (tini + UID mismatch + sessions)
- /goal judge over-continues exploratory goals unless the assistant explicitly says the goal is complete
- /goal auto-continuation can be amplified by preflight compression/session split and resurrect stale task state
- Dashboard infinite reload loop in loopback mode — GET /api/auth/me returns 401 on every page load
- [Bug]: Provider/LLM switch leaves stale encrypted_content causing 400 errors on Telegram sessions
- [Bug]: Infinite reload loop / React state loop on Sessions tab (Firefox + Chrome) — repeated 401 on /api/auth/me (v0.15.0)
- show_reasoning should work independently of streaming in CLI mode
- Feature Request: Strip reasoning/<think> blocks from TTS preprocessing
- mcp add / mcp test raise NameError when mcp package not installed
- v0.14.0 dashboard breaks behind reverse proxies — two regressions
- Skills hub creates empty category directories when no skills installed
- [Bug]: Custom endpoint: ChatCompletions returns content, but Hermes treats response as empty (v0.14.0)
- fix: atomic_replace() fails with EXDEV when HERMES_HOME is a cross-filesystem symlink
- fix(gateway): Feishu session cancellation orphans session guard, permanently blocking messages
- Custom endpoint pricing can overestimate Crof qwen3.5-9b cost by 1,000,000x
- MCP OAuth callback: module-level port global causes port collisions and structural weaknesses vs upstream
- Bug: send_message tool bypasses validate_media_delivery_path security check
- Proposal: Add Mnemosyne to official memory provider documentation
- feat(swarm): support custom verifier/synthesizer body + skills
- Template conversion failed
- Error occurred in the operation of the agent node in the workflow.
- PubSub client overrides Sentinel client when REDIS_USE_SENTINEL is enabled
- Frontend description of the Retrieval node output does not match the actual output
- JSON type input var raise Intenal server error
- cannot extract elements from a scalar
- 负载均衡 为模型配置多组凭据,并自动调用,此功能无法选择
- add models is error
- panic: could not create filter
- Persist partially generated messages when /chat-messages/:task_id/stop is called
- MCP server connection fails with 403 — request never leaves Dify (SSRF proxy suspected)
- Support durable async execution backends for long-running workflow steps
- [Xiaomi MiMo] Credentials validation fails with 400 "Not supported model mimo-v2-flash" when using Token Plan endpoint (v0.0.7)
- After clicking preview on a parent-child segmented knowledge base, it shows 0 chunks
- Retrieval score differs between UI upload (.docx) and API upload (.txt) despite identical chunk content and embedding model
- gemini cli crash again
- Xbox gift card code damage
- Damage caused by the gemini cli crash
- ioctl(2) failed, EBADF (Bad File Descriptor)
- Feat: Support Bun as an alternative runtime/package manager for updates and extensions
- fatal error again!!!!
- ioctl error
- Critical Crash: ioctl(2) failed, EBADF in ShellExecutionService.resizePty
- ioctl(2) failed, EBADF
- v0.44.0 Regression: Critical crash with ioctl(2) failed, EBADF during PTY resize
- Crash on startup: ioctl(2) failed, EBADF in UnixTerminal.resize
- Crash: `ioctl(2) failed, EBADF` in `node-pty` during PTY resize on macOS
- Gemini CLI crashes with `ioctl(2) failed, EBADF` in `node-pty` during `resizePty`
- Remote Role
- ERROR ioctl(2) failed, EBADF /home/mich
- RangeError: Maximum call stack size exceeded
- EBADF Error during folder creationg broke session and terminal glitches
- MAIP / Gargoub Project - Mediterania - North Coast
- Gemini cli crash again in this morning
- ERROR ioctl(2) failed, EBADF
- Verified node install fails — Checksum verification failed (Cloud)
- The extended debugging key did not arrive during registration.
- CollaborationPane unmounts collaboration store on single-user instances, causing permanent "No network connection" state
- Workflow cannot be saved when the name contains "->" (Potentially malicious string)
- automation does not work and does not show an error
- Raj Ai Automation
- Default Data Loader: DOMMatrix is not defined error
- Feature: Per-node execution timestamp overlay on canvas during workflow run
- AI Agent + Vertex `gemini-3.5-flash`: 400 "missing thought_signature" on sequential multi-turn tool calls (post-#24982)
- PDF Loader in Pinecone Vector Store fails due to pdf-parse version conflict (v2 not supported)
- emailReadImap: add UID deduplication, batch size cap, and numeric uid enforcement
- Manual node execution fails with "Could not find a node" when autosave is disabled (N8N_WORKFLOWS_AUTOSAVE_DISABLED)
- Schedule Trigger stopped firing — workflow Published & active, manual executions succeed, no automated fires for 2+ hours
- [MCP SDK] create_workflow_from_code intermittently returns HTTP 500, often as a false negative (workflow persists anyway, causing duplicates on retry)
- Credential-load wedge: workflows using googleApi/jwtAuth credentials silently fail to execute after key rotation
- Google Sheets Trigger every minute is not working manual Execute is working sent email
- [BUG] Plugin marketplace MCP connector remains stuck "still connecting" when mcp-remote requires OAuth
- [redacted at user request]
- Opus 4.7 behavioral regression: loaded instruction-following discipline degraded in recent Claude Code/Cowork updates
- [BUG] Tailscale via Homebrew CLI + Mac App Store GUI, both Macs on macOS, Cowork blocked by VPN detector despite Tailscale being a mesh VPN with no traffic interception
- stopShellPty on tab switch kills active sessions (exit 143) — regression in May 27 build
- [BUG] Long URLs are broken into multiple lines and become unclickable in terminal output
- [BUG] claude rm/stop/reap SIGKILLs background session tree without SIGTERM grace, orphaning git index.lock and similar
- [BUG] Default git workflow in the system prompt was pushed without context or consent
- [MODEL] Inconsistent output quality / Ignoring instructions (overfitting and inappropriate repetition of Korean vocabulary)
- You've hit your weekly limit · resets May 31 at 5pm (Asia/Shanghai)
- Paid yearly subscription silently downgraded to Free with no user action
- [Regression v2.1.153] Plugin bash hooks fail with "echo: write error: Permission denied" on Windows (claude-mem, shell: "bash")
- [BUG] Connector toggles in conversation are not clickable — must click text label instead
- [remote-control] Input from mobile app/browser not reaching host session — output works fine
- Model fails to read/reference CLAUDE.md contents despite being loaded in context
- [BUG] Claude Desktop reinstall destroys Code chat history (transcripts + Recents) while regular Chat history, project files, and memory all survive
- Bypass mode clamps to Accept Edits even with the toggle ON (Claude Code Desktop 1.9255.2 / CC 2.1.149)
- [BUG] TUI input freezes randomly mid-typing — entire prompt becomes unresponsive for minutes
- [BUG] Cowork downloads Linux ELF binary instead of macOS binary on macOS Sonoma 14.8.7 — exit code 132 (SIGILL) on every session
- [Feature Request] Persistent project memory — sessions forget everything on close, forcing users to keep many sessions open
- [Bug] Thread context stale after sleep/resume, returns outdated date and calendar data
- [FEATURE] Add context window usage indicator and warning before auto-compaction
- [BUG] Dictation error: Invalid character in header content ["x-config-keyterms"] on Windows
- [Bug] Anthropic API Error: Server rate limiting despite normal usage
- Does delegating work to `claude -p` subprocesses reduce context accumulation in the parent session?
- [BUG] Claude Code hangs on M1 Mac when terminal says "opening browser to sign in" and browser opens
- [BUG] Claude_Preview MCP preview_start spawns dev server with main-repo cwd instead of session's worktree cwd
- [Bug] Anthropic API Error: Server rate limiting during request execution
- [Bug] Anthropic API Error: Server rate limiting on concurrent requests
- [Bug] Ultraplan ready notification fires before cloud agent completes execution
- [BUG] API 500 ERROR ALL THROUGHOUT THE DAY
- [BUG] Cowork: Live Artifacts folder path changed in 1.9255.2, no automatic migration from Documents\Claude\Artifacts
- [Bug] Auto-compact never triggers despite statusline reporting "100% context used" (v2.1.153, Max sub, 200K mode)
- [BUG] [Desktop / macOS] 'Open in → New Window' detached session: font renders smaller than main, no per-window controls, Cmd+/Cmd- keystrokes routed to main window instead
- Feature request: option to switch between classic and new minimal UI
- [Feature Request] Show timestamps for each message
- [BUG] Terminal corruption when permission prompt appears while navigating Agent Teams agent selection menu
- [FEATURE] Allow users to customize the background color of the Claude desktop app beyond the current light/dark theme presets.
- [BUG] Statusline not displaying on Windows [fixed]
- Background agent UI Stop button is a no-op for stuck agents — process keeps consuming tokens
- Background agents silently die on session pause/resume — no completion notification, no work recovery
- Add option to hide email address from welcome banner
- [BUG] SSH Remote: `projects` field in remote ~/.claude.json becomes null after desktop restart — jsonl files intact, UI shows 'No messages yet' for every session
- [Bug] Claude Code not applying fixes despite claiming to complete tasks
- billing is unfair and poorly documented
- [BUG] Claude Code on the web: declared plugins inactive on first session, require restart to fully load
- [BUG] Restore from archive deleted sessions instead of restoring them
- [BUG] M365 connector fails with AADSTS50011 in Cowork — localhost vs 127.0.0.1 redirect URI mismatch
- claude agents: workflow slash-commands missing from dispatch-input completion (regression-adjacent to #61424)
- Claude Desktop's Info.plist missing TCC usage strings, blocks all EventKit-based MCP servers
- False-positive safety blocks on self-administered governance amendments — request for owner-authority mode for verified professional users
- [BUG] Stop pushing "AUTO"-mode
- [DOCS] Plugin marketplace guide omits `skipLfs` option for git-based sources
- [DOCS] MCP docs omit combined startup notification for MCP server and connector authentication
- [DOCS] Agent view docs omit macOS Privacy & Security identity for background agents
- [DOCS] Npm update docs do not explain release-channel behavior for `claude update`
- [DOCS] Agent SDK docs omit `subagent_type: "claude"` worktree and output persistence behavior
- [DOCS] Background session docs omit `$CLAUDE_JOB_DIR` temp-file behavior
- [FR] mask env-var values in 'claude mcp get <server>' output
- [FR] subagent worktrees should not inherit stale local 'user.email' from prior dispatches
- [BUG] Windows: Grep tool leaks rg.exe + conhost.exe processes (~2000 zombies / 14 GB RAM in long sessions)
- [BUG] Stats dashboard "Peak hour" appears off by one hour
- [BUG] Diff highlight (teal SGR background) bleeds past changed text in 2.1.150–2.1.153
- [FEATURE] confirm before deleting session
- Plugin PostToolUse hooks still silently skip in Claude Desktop / Cowork (re-filing closed #51904)
- /code-review skill: silent fallback to main...HEAD reviews other people's commits, and JSON-only output is hard to read
- Monitor tool doesn't source the shell snapshot like Bash does; PATH-dependent tools (jq, sleep, etc.) fail in Monitor commands on macOS/Nix
- [Bug] Long input lines truncated with ellipsis while typing instead of wrapping in terminal UI
- [FEATURE] VS Code extension: Render submitted user messages as Markdown in chat
- OSC 52 copy from Claude TUI doesn't reach clipboard inside tmux (regression in 2.1.146–2.1.153)
- [BUG] RemoteTrigger create/update returns HTTP 400 with circular error: "event_type is required" / "unknown field event_type"
- [BUG] Option to hide or minimize the built-in "status footer" (multi-line debug/cost panel) [re-raise of #31475]
- [Bug] Feedback submissions being closed without review or action
- [FEATURE] Word-jump cursor navigation in Chat input (option+arrow / bindable actions)
- [FEATURE] ! shell mode: filesystem tab completion
- [BUG] API Error: Usage credits required for 1M context
- claude agents: OSC 52 clipboard emission broken in tmux (regression in 2.1.146–2.1.153)
- CLI crashes on macOS 15 M3 - exit code 1
- [FEATURE] Support Cmd+V image paste from clipboard
- [FEATURE] Enhance claude.ai M365 connector to support MS Planner
- [BUG] Slash command autocomplete hijacks pasted absolute file paths starting with /
- PreToolUse hook `if` filter false-positives on complex Bash commands
- [BUG] Diff panel hangs/whites out
- Feature Request: Support drag-and-drop for binary documents (.wps, .doc, .docx, .xlsx, .pdf) in VS Code extension
- [BUG] activation of 1M context in VSCode
- [FEATURE] Support i18n / language localization for built-in slash command outputs
- Ctrl+V para colar imagens deixou de funcionar no CLI (Windows, PowerShell)
- [FEATURE] Please add Norwegian (Bokmål/Nynorsk) language support to the Claude Code interface
- [BUG] OTel log events (claude_code.user_prompt, api_request_body, tool_decision, hook_execution_complete) emitted with empty trace_id/span_id while sibling spans correlate correctly
- [BUG] Cowork crashes on every message, no VM logs generated, missing AppData\Roaming\Claude
- [FEATURE] first-class session handoff + per-session token budgets for unattended runs
- [FEATURE] Smart paste: convert clipboard code to file reference chips (like Cursor)
- [Feature Request] Restore chat pin functionality to title chat submenu
- [BUG] SIGILL issues with version 2.1.153
- [BUG] Cowork plugin upload fails with generic "Plugin validation failed" when a `description` field in any SKILL.md frontmatter contains angle brackets (`<…>`)
- [BUG] Desktop App 2.1.144+: startup scanner deletes cliSessionId from claude-code-sessions local files on every launch — session not found on disk
- [Feature Request] Add keyboard shortcut to copy last message with proper formatting
- [MODEL] Opus 4.7 not 1M
- Allow naming/renaming background agents in `claude agents` view
- Stale worktrees in .claude/worktrees/ are never cleaned up, consuming massive disk space
- Agent worktrees are never cleaned up, silently consuming disk space
- Subagent worktrees not auto-cleaned when reviewer writes scratch files
- [Bug] Skill initialization hangs for extended duration in Plan Mode
- Claude Desktop writes malformed registry Run entry (nested escaped quotes) - crashes Windows Task Manager and other Run-key parsers
- IME candidate window shows at bottom-right corner instead of caret position (Windows CMD)
- [BUG] Pressing 'Escape' doesn't close the /BTW conversation when the main conversation is asking for approval
- [BUG] Opus 4.7 (1M) intermittently emits empty-string values for tool_use.input fields, killing the session
- FleetView agent UI shows "running" with incrementing elapsed time after agent has returned
- /doctor flags context-scoped cmd+c binding as macOS conflict (false positive)
- [BUG] Text Rendering in Elvish
- Desktop app: Bypass Permissions mode flips to Accept Edits on first prompt (M5 / macOS 26.5)
- [Workaround] Date-Weekday Verification Hook — Prevents Claude from writing wrong weekdays
- [BUG] Claude Code create c:/memfs directory without asking me.
- [BUG] Claude Code's Bash execution waits forever with no processes running
- [BUG] usage stays stuck waiting for 5 hr limit after upgrading to premium seat in team plan
- [Workflow tool] resume cache is unreachable for nontrivial workflows because LLM dispatchers can't transcribe args byte-exactly
- Code review (Preview): "Add a repository" shows no results for private GitHub org repos
- [BUG] /context commands blows up context
- [Feature Request] Add precache expiry hook to enable proactive compaction before token eviction
- [BUG] Context indicator shows 0% at session start despite ~20K+ tokens already loaded
- [Feature Request] Add semantic search for --resume session history
- [Feature Request] Add session search, tagging, and filtering capabilities
- [BUG] Cowork Dispatch reports "desktop not available" on Windows 11 while standard Cowork works normally
- [Bug] Claude Code provides incorrect suggestions with high confidence despite errors
- defaultMode: acceptEdits silently overrides per-path permissions.ask rules for Write/Edit
- [FEATUR configurable tip interval (e.g. tipIntervalSeconds: 30 in settings)E]
- Plugin marketplace fails to load: schema rejects 'displayName' key (v2.1.153)
- claude agents: in-session copy uses broken OSC 52 path while overview correctly uses tmux buffer
- [BUG] Plugin agent descriptions (and custom agents) load unconditionally into context — no parity with disable-model-invocation for skills
- Crashed ultrareview consumed a free credit despite producing zero findings
- [Bug] Character rendering issue - invisible or missing text display
- [BUG] Cowork: processo Claude Code encerra com código 3 — .claude.json não contém token de autenticação (Windows 11 25H2)
- [BUG] 2.1.153 silently discards tools/list response from rmcp 0.12.0 HTTP MCP server (works in 2.1.152, wire-identical handshake)
- VS Code extension: option to auto-resume last session when reopening a workspace folder
- [Bug] Conversation continuation failure
- [BUG] Cowork crashes every time I start a new chat or attempt to continue an existing one in any project. The error displayed is: "Claude Code è andato in crash
- [Bug] Unannounced quota changes
- Native update/install fails with 'socket connection was closed unexpectedly' behind proxy — undici TLS incompatibility
- [BUG] Session name reverting after manual change
- [BUG] 非正常思考,上下文过长时,一直显示思考,点击interrupt按钮失效
- Honor `tools:` frontmatter when an agent is invoked via `@mention` — strip `Task` only when the agent did not declare it
- macOS TCC popup still recurring on v2.1.153 — "2.1.153" would like to access data from other apps
- Claude Code leaks pty handles — exhausts pseudo-terminals on macOS after long session
- [Bug] Agent fails to execute or respond to user input
- [BUG] Persistent "Expecting value: line 1 column 1 (char 0)" JSON parse error after tool execution
- [Feature Request] Implement proactive unit test coverage recommendations for recurring bugs
- VS Code panel lacks status line + terminal lacks image paste in Codespaces, forcing a tradeoff
- `/powerup` only shows ~10 lessons — allow viewing the full catalog
- [Bug] Context contamination after auto-compact with unrelated email draft of Tejo/Sado Basin
- [Bug] VSCode terminal output displays corrupted text with garbled symbols
- [Feature Request] Add LaTeX/KaTeX math rendering to TUI
- [Bug] Sub-agent PR review results not validated by orchestrating agent
- Subagents on Pro 1M tier: trivial probes pass, real workloads fail at first tool call (probe-vs-workload divergence)
- Path-scoped rules and subdirectory CLAUDE.md not loaded when creating new files matching the pattern
- AskUserQuestion: cancelling during extended thinking poisons the whole session with 400 'thinking blocks cannot be modified' (2.1.153); concurrent prompts overwrite each other
- Ideas Missing from Claude Cowork Menu (Windows)
- [BUG_BOUNTY_SAFE_POC_2026] Prompt Injection RCE Test - Command Execution Proof
- [BUG] Cowork scheduled task: execution history row not showing after successful run
- Resuming an extended-thinking session fails permanently with 400 "thinking blocks cannot be modified" (transcript stores thinking text as empty but keeps signature)
- [Bug] Plugin-registered CwdChanged and FileChanged hooks don't fire (settings.json works) — v2.1.153
- Auto-archive on PR merge / branch delete — clarify autoArchiveSessions semantics or add dedicated opt-out
- `claude mcp add` echoes Authorization header value verbatim to stdout, leaks bearer tokens to terminal and session transcripts
- [BUG] Bug report — /insights skill, Claude Code The /insights skill outputs a malformed file path.
- Plugin slash commands render with '*'-inline format instead of two-column, despite matching official plugin shape
- [Bug] Unexpected long text generation without user input or goal
- [Bug] Thinking blocks causing task progression blocked without user modification
- [BUG] (Critical!) contamination by an unknown session simirlar to the report => [Bug] Context contamination after auto-compact with unrelated email draft of Tejo/Sado Basin #63137
- [Critical] Opus 4.7 Korean output degeneration — Korean grammar itself collapses in long contexts
- [BUG] Title: Autocompact buffer persists across /clear — wastes tokens for irrelevant old context
- [Bug] Auto-Compact loses user input before processing in conversation history
- Feature: per-invocation effort parameter + runtime session-config introspection for skills
- Auto-mode classifier mislabels Azure DevOps vote -5 as "Reject" when denying PR vote actions
- [BUG] Claude Desktop and Claude Code CLI never re-register MCP tools after OAuth 2.1 handshake on a remote HTTP server
- [BUG] Workspace file tags leak across sessions
- [BUG] Ink renderer crashes on Windows 11 build 26200 (Canary) duplicate banners, terminal mode leaks, mid-operation aborts
- [BUG] Claude Code Desktop issue
- PTY master fd leak in Claude desktop app exhausts macOS kern.tty.ptmx_max after ~2-3 days
- [BUG] Claude Code — Session Management after Unexpected Interruption
- [Windows] Cowork OpenTelemetry exporter does not initialize - zero events emitted to any destination, including loopback
- [Bug] Opus 4.7: 400 `thinking blocks ... cannot be modified` on long extended-thinking sessions, triggered by history-altering events (scheduled prompts / parallel tool-call cancellation)
- [BUG] API Error: Server is temporarily limiting requests (not your usage limit) · Rate limited
- Multi-plugin custom marketplace: only first plugin registered in installed_plugins.json, skills don't load
- [BUG] Git push through the SDK's git proxy fan-outs into ~500 GitHub REST API calls, exhausting the 5,000/hour budget after a handful of pushes
- [BUG] Claude took liberties it really shouldn't with my global config
- [BUG] Agent window focus lost after navigating with arrow keys, causing scroll deadlock
- [BUG] `--model` flag silently ignored in interactive sessions (works in `--print` only)
- [BUG] Dispatch permanently shows "desktop appears offline" on Windows 11 - never worked on first use
- feat: support per-command enableWeakerNetworkIsolation as safer alternative to dangerouslyDisableSandbox
- /code-review outputs a raw JSON array instead of readable findings
- [BUG] Cowork — Additional allowed domains ignored on Team plan; same domain works on Pro plan
- Haiku
- [Bug] False positive blocking beneficial outcomes in tool execution
- 3P Bedrock SSO: credentials silently expire without triggering re-auth on day 2+
- CLAUDE_AUTOCOMPACT_PCT_OVERRIDE in settings.json env block silently ignored by autocompact logic
- Auto-compaction deletes main session JSONL before verifying summary completion, causing data loss
- [Bug] Claude Code not executing stated actions or producing expected results
- [FEATURE] Deferred Messages — Queue Input for End of Turn
- [BUG] Up/Down arrows in input box navigate history instead of moving cursor — regression in 2.1.149+
- Cancelling a parallel tool-call batch corrupts thinking blocks -> 400 "thinking blocks cannot be modified" permanently wedges the session
- Claude Code caused data loss, then contradicted itself about recovery (two incidents, one session)
- [Bug] Unclear error messages from Claude Code CLI
- [Bug] Agent tool rejecting due to context size limit exceeded
- claude agents: daemon and bg-spare processes spin at ~100% CPU when idle
- [BUG] Compaction fails with "context window limit" error even when context usage is low (e.g., 20%) — regression in v2.1.153
- Remote Control entitlement lost after May 27-28 incident — `Error: Remote Control is not yet enabled for your account` on active Max subscription
- PreToolUse hook exit code 2 does not block Write tool
- [Bug] Thinking blocks in latest assistant message are immutable
- GUI: dispatch file:// and custom-scheme clicks to OS shell handler
- Show current model in statusLine by default
- [Bug] Agent console becomes unresponsive to keyboard input after multiple agents initialized
- [FEATURE] PreToolUse hooks should have a way of updating the environment
- [Bug] Unable to start or use Claude Code CLI
- [BUG] Repository not visible in Claude Code web repo picker
- Session permanently wedged on 400 "thinking blocks cannot be modified" after parallel tool_results
- [Bug] @ autocomplete loses sibling repos after a file edit in multi-repo workspace
- Unclear error message when creating sub-agent without authentication
- [Bug] Anthropic API errors causing frequent failures and high token usage
- [BUG] @ mention file picker only shows packages, not individual files (desktop app - Code tab)
- [Bug] TUI panel footer remains sticky and consumes excessive terminal space
- PR-status polling exhausts GitHub GraphQL rate limit on repos with many open PRs
- [BUG] Windows: welcome panel not shown in some project folders (2.1.153)
- [Bug] Anthropic API Error: thinking blocks corrupted during context compaction with extended thinking enabled
- API 400 "thinking blocks cannot be modified" permanently bricks session during agent activation (interleaved thinking + tool use)
- Right-click Copy copies the whole message instead of the selection; pasted text retains dark background
- Mid-session model switch corrupts conversation when extended thinking is enabled (API 400: 'thinking blocks cannot be modified')
- [BUG] Markdown file links in chat output do not open files when clicked (VS Code extension)
- Stuck retry loop: `400 thinking blocks cannot be modified` on large interleaved-thinking turns using AskUserQuestion
- [FEATURE] Prompt user for approval before auto-compaction proceeds
- Custom MCP connectors not attachable to scheduled routines — no UUID discovery path
- [BUG] Claude in Chrome — Navigation blocked for teams.cloud.microsoft and outlook.cloud.microsoft after Microsoft domain migration**
- [BUG] Claude Desktop — Personal plugins panel renders list but is entirely non-interactive (macOS, v1.9255.2)
- [Bug] error when using Workflows
- [BUG] Persistent "update available" notification despite being on latest version
- [BUG] Sweep Agent from /code-review never completes
- [Bug] Tool calls not executing or returning results
- [FEATURE] Cloud-synced memory and settings across machines
- [Bug] Terminal UI freezes when Ctrl+O view exits during interactive prompt in plan mode
- Continuous api errors when using claude code with Opus 4.7 with thinking on low
- [Feature Request] Add support for installing and using previous Claude Code versions
- [Bug] Extended Thinking: Summarized thinking blocks fail signature validation when resent to API
- [Bug] Anthropic API Error: 'thinking' blocks cannot be modified
- [Bug] Anthropic API Error: Thinking blocks cannot be modified with extended thinking mode
- Feature request: Lazy/on-demand MCP server connections
- [Bug] Tool Arguments Parsed as String Instead of Object
- [Bug] Anthropic API Error: Insufficient context provided
- [Bug] Claude Opus occasionally uses moskovian(russian) orthography instead of Ukrainian in system-prompted responses
- Opus 4.8: backgrounded task completions (subagents AND Bash) crash with 400 "thinking blocks cannot be modified"
- [Bug] Opus 4.7 fabricates stable preferences ("my default") to rationalize arbitrary choices when challenged
- [Bug] Unable to update Claude Code CLI
- [BUG] Desktop app: /remote-control mints link + connects bridge (main.log) but in-chat link/QR panel never renders
- Feature: sessionColor and sessionName in .claude/settings.json
- [BUG] Anthropic API error: thinking blocks
- [FEATURE] Support Remote MCPs in Cowork as in Claude Code
- [Bug] Anthropic API Error: 400 Bad Request with Redacted Thinking - 0 4.7 & 4.8
- [Bug] Anthropic API Error: Cannot modify thinking blocks from different model versions
- Interleaved thinking + multi-tool turn corrupts thinking block (text blanked, signature kept) → permanent 400 'blocks must remain as they were'
- [BUG] Mode/permission changes mid-tool-loop (effortLevel: xhigh) poisons entire session
- Session failure log: Opus 4.6 ignores its own rules for an entire session
- [BUG] "400 Guardrail was enabled" error when using Claude Opus 4.8 with AWS Bedrock
- [Feature Request] Add subagent approach selection option to avoid accidental feedback
- Persistent 400 'thinking blocks in the latest assistant message cannot be modified' — interleaved thinking persisted with empty text + signature bricks sessions
- [BUG] DesktopvsApp
- [BUG] Opus 4.7 cache hit rate collapse after May 27 incident — Messages 1.1k→88.9k in 9 minutes, $630/session
- [Bug] Anthropic API Error: Invalid thinking block format
- [BUG] FUCK CLAUDE
- Opus 4.8 extended thinking: Stop hook block re-entry corrupts thinking blocks → 400
- [Bug] 4.8 Fails when accessing previous model history
- [Bug] Unintended File Modifications During Execution
- [DOCS] Model configuration docs omit lean system prompt default scope and model exceptions
- Add "Always allow globally" option to permission prompts
- Server-side model upgrade (Opus 4.7→4.8) wedges in-flight sessions with `thinking blocks cannot be modified` 400
- [DOCS] AskUserQuestion docs missing multiple-choice prompt decision threshold
- [DOCS] Agent view docs omit shell-command background session launch syntax
- [DOCS] Agent view dispatch input docs incorrectly imply `/logout` dispatches as a prompt
- [DOCS] Claude in Chrome docs omit connected-browser selection behavior
- [DOCS] Plugin docs omit `defaultEnabled: false` for opt-in plugins
- Feature Request: Customizable chat text colors for user and assistant messages
- [DOCS] `/plugin` Discover tab docs omit directory-based suggested plugin pins
- VSCode Chrome integration silently fails: 3 distinct bugs
- [DOCS] MCP stdio docs omit session environment variables
- [Bug] Anthropic API error on second request within session with Claude Opus 4.8
- Cowork emits a blank session "index" handoff on focus when a CLI session is paused awaiting input
- [DOCS] MCP docs omit `claude mcp list/get` pending-approval output for unapproved project servers
- [BUG] /compact fails with 400 error when last assistant turn contains thinking blocks
- [DOCS] `/claude-api` docs omit Opus 4.8 migration guidance
- [DOCS] Fast mode docs still recommend deprecated Opus 4.6 override variable
- [DOCS] Bash tool docs omit `$TMPDIR` consistency across sandboxed and unsandboxed commands
- [Bug] Anthropic API Error: 400 Bad Request on Extended Thinking
- [DOCS] Background session docs omit worktree-isolation behavior for spawned subagents
- Built-in mechanistic self-verification of verifiable claims (symmetric to the auto permission gate)
- [DOCS] Worktree docs do not clarify `worktree.baseRef: "head"` inside linked worktrees
- [BUG] Excessive RAM usage with multiple parallel chats (~10 sessions → 30 GB memory pressure, macOS OOM)
- [DOCS] Managed MCP policy docs omit invalid `allowedMcpServers`/`deniedMcpServers` entry behavior
- [DOCS] Effort docs omit `CLAUDE_CODE_ALWAYS_ENABLE_EFFORT` unsupported-model behavior
- Regression (2.1.147–2.1.150?): resuming an extended-thinking session after a CC update/model-switch → unrecoverable 400, session bricked
- [DOCS] Windows updater docs omit `claude.exe` in-use recovery guidance
- [DOCS] VS Code auto mode docs still tie mode-picker visibility to bypass-permissions setting
- [DOCS] MCP docs omit `/mcp` tool list and detail rendering behavior
- [DOCS] Fine-grained tool streaming docs still describe provider opt-in behavior
- bypassPermissions: session startup reads flat pref, GUI toggle writes per-account pref — they never sync
- [BUG] Claude Desktop Code tab causes disk write limit violation — 8.5GB in 11 min, macOS kills app (M5, v1.9659.1)
- Ultrareview v2.1.96: docs describe /tasks command + claude ultrareview --json subcommand that don't exist; findings hard to read after completion
- I'd be happy to help create a GitHub issue title, but I don't see the error message in your message. Could you please share the specific error you're encountering? That way I can generate an accurate and descriptive issue title for you.
- [BUG] Claude in Chrome `file_upload` rejects all scheduled-task sessions with misleading error (real cause: INVALID_SESSION)
- Extended thinking: signed thinking block 'cannot be modified' (400) permanently wedges session
- RTL text support for Hebrew (and Arabic) in Claude Code
- [Bug] Random errors occurring across multiple operations