文件历史

提交图

9 次代码提交

作者 SHA1 备注 提交日期
Simon Willison 5a2e0a4c54 --hide-reasoning and hide_reasoning=True parameters (#1442)
* Rename --no-reasoning flag to --hide-reasoning
* hide_reasoning= Prompt parameter, plus docs
* OpenAI plugin now obeys prompt.hide_reasoning
2026-05-12 09:24:51 -07:00
Simon Willison 6e12258c0b Send prior assistant text as plain string in OpenAI Responses input
When building the `input` list for the OpenAI Responses API from prior
conversation turns, an assistant text-only turn was being serialized as:

    {"role": "assistant",
     "content": [{"type": "output_text", "text": "..."}]}

The openai-python SDK's EasyInputMessage shape uses a plain string for
this case, matching what a direct OpenAI Responses call would send. Use
the same shape so our history matches the SDK exactly, and add tests
covering both _build_responses_input and a two-turn response.reply()
flow.
2026-05-12 08:19:17 -07:00
Simon Willison 3e9e3a30ea Request reasoning summary auto for OpenAI models 2026-05-11 23:06:52 -07:00
Simon Willison 98e651075f Register more OpenAI models using Responses API 2026-05-11 22:33:28 -07:00
Simon Willison 2297a2aab0 Fix for ruff 2026-05-11 21:37:47 -07:00
Simon Willison 2838388c31 Ran Black 2026-05-11 21:18:40 -07:00
Claude 1a56805ceb Verify tool calls during reasoning work end-to-end
The previous commit wired up encrypted_content round-trip but only
tested that the data flows through correctly on a single tool round-
trip. This adds a multi-turn cassette test that proves the full
interleaved-reasoning capability:

- Each turn produces fresh reasoning_tokens (not just the first)
- Every prior reasoning block is round-tripped on every subsequent
  turn (the Nth turn echoes >= N-1 reasoning items)
- ReasoningParts persisted on the assistant messages carry the same
  encrypted_content + id that gets sent back on the wire

The puzzle is shaped so the model can't parallelize tool calls -
each db_lookup result tells it the next key to use, forcing the
model to think between calls. The recorded 4-turn chain shows
reasoning_tokens of 45/98/196/17 across turns with reasoning items
accumulating in every outgoing input.

This is the GPT-5-class capability that Chat Completions can't
deliver because it discards reasoning between turns.
2026-05-06 04:14:04 +00:00
Claude c7464eeb97 Round-trip encrypted reasoning across tool calls
When the Responses API returns a reasoning item alongside function
calls, capture its opaque id + encrypted_content as provider_metadata
on the resulting ReasoningPart. _build_responses_input already echoed
that metadata back as a reasoning input item on the next turn - now
the output side actually populates it.

This preserves the model's hidden chain of thought across the tool
round-trip. Without it, GPT-5-class models silently lose ~3% on
SWE-bench (per OpenAI) when used with tools.

Adds a dedicated VCR test that asserts the encrypted_content captured
on the first turn appears verbatim in the second turn's outgoing
request body.
2026-05-06 03:15:21 +00:00
Claude 3c747c8b9f Route gpt-5.5 through the /v1/responses endpoint
Adds Responses and AsyncResponses classes that drive the OpenAI
/v1/responses endpoint. The existing Chat / AsyncChat classes are
unchanged because other plugins import them.

gpt-5.5 (and gpt-5.5-2026-04-23) is now registered against Responses
by default. Pass `-o chat_completions 1` to fall back to the older
/v1/chat/completions code path.

This is feature parity with the Chat path (text, tools, streaming,
schema, reasoning_effort, verbosity, attachments, system prompts).
Interleaved reasoning across tool round-trips is not exercised yet -
encrypted reasoning items are accepted on the input side, but the
plugin doesn't yet stash them on outgoing ReasoningParts.
2026-05-06 01:57:09 +00:00