Skip to content

Commit 616f24b

Browse files
docs(tests): the gateway thinking-block bug is fixed upstream; the env var is now a cost choice
Proxy upgraded to LiteLLM v1.101.0-dev.2 on 2026-09-05 (digest-pinned in the proxy stack, not here). The upgrade path had a wrinkle worth recording: v1.99.1 fixed the thinking-block adapter crash but regressed bearer-token Bedrock auth ("NoneType has no attribute access_key", upstream #38579); the dev.2 build carries both fixes, and a forced thinking block now streams cleanly through /v1/messages (verified with an explicit thinking budget, content blocks {thinking, text}, no stream error). MAX_THINKING_TOKENS=0 stays: thinking adds nothing to a protocol-following check and costs tokens. The comment now says that is a cost choice, not a workaround.
1 parent cb29794 commit 616f24b

1 file changed

Lines changed: 8 additions & 7 deletions

File tree

‎packages/studyloop/tests/live/test_wind_down_transcripts.py‎

Lines changed: 8 additions & 7 deletions
Original file line numberDiff line numberDiff line change
@@ -214,13 +214,14 @@ def _child_env(home: Path, config_path: Path) -> dict[str, str]:
214214
"STUDYLOOP_PLANS_DIR": str(home / "plans"),
215215
"DISABLE_AUTOUPDATER": "1",
216216
"CLAUDE_CODE_DISABLE_NONESSENTIAL_TRAFFIC": "1",
217-
# The gateway's anthropic-passthrough adapter (LiteLLM, Bedrock invoke
218-
# path) dies on extended-thinking blocks: "API Error: Content block is
219-
# not a text block", reproduced deterministically on prompts long
220-
# enough to trigger thinking and absent on short ones (proxy logs,
221-
# 2026-09-04). Thinking adds nothing to a protocol-following check, so
222-
# it is off rather than worked around. Remove once the proxy is
223-
# upgraded past the adapter bug.
217+
# Thinking is off because it adds nothing to a protocol-following
218+
# check and costs tokens. It USED to be mandatory: LiteLLM <= 1.99.x
219+
# died on extended-thinking blocks over the Bedrock invoke path
220+
# ("API Error: Content block is not a text block", proxy logs
221+
# 2026-09-04). Fixed upstream (BerriAI/litellm PR #33315 + the
222+
# bearer-token repair in #39166); the proxy runs v1.101.0-dev.2 as of
223+
# 2026-09-05 and a forced thinking block streams cleanly through
224+
# /v1/messages — verified with an explicit thinking-budget request.
224225
"MAX_THINKING_TOKENS": "0",
225226
"TERM": "dumb",
226227
}

0 commit comments

Comments
 (0)