You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
docs(integrations): recommend anthropic_messages for Claude through an LLM gateway (#874)
* docs(integrations): add Claude gateway caching and thinking notes
Signed-off-by: Elyas Mehtabuddin <emehtabuddin@nvidia.com>
* docs(integrations): clarify when each Claude gateway setup applies
Signed-off-by: Elyas Mehtabuddin <emehtabuddin@nvidia.com>
* docs(integrations): estimate Claude cache costs and fix the thinking and key setup steps
Signed-off-by: Elyas Mehtabuddin <emehtabuddin@nvidia.com>
* docs(integrations): note when omit_body_fields holds and use Anthropic's published cache prices
Signed-off-by: Elyas Mehtabuddin <emehtabuddin@nvidia.com>
* docs(integrations): document forwarding one gateway key to OpenAI and Anthropic clients
Signed-off-by: Elyas Mehtabuddin <emehtabuddin@nvidia.com>
* docs(integrations): say only standalone switchyard-server forwards keys
Signed-off-by: Elyas Mehtabuddin <emehtabuddin@nvidia.com>
* docs(integrations): lead the Claude gateway sections with the format to use
Signed-off-by: Elyas Mehtabuddin <emehtabuddin@nvidia.com>
---------
Signed-off-by: Elyas Mehtabuddin <emehtabuddin@nvidia.com>
Copy file name to clipboardExpand all lines: docs/reference/toml_schema.md
+1Lines changed: 1 addition & 0 deletions
Display the source diff
Display the rich diff
Original file line number
Diff line number
Diff line change
@@ -149,6 +149,7 @@ such clients.
149
149
|`llm_client`| Yes | — | Key under `[llm_clients]`. |
150
150
|`system_prompt`| No | unset | System prompt prepended when this target serves a completion. |
151
151
|`extra_body`| No |`{}`| Values merged into the upstream request when the request does not already set that key. |
152
+
|`omit_body_fields`| No |`[]`| Top-level fields removed from every request body that Switchyard sends to this target. Switchyard removes them after it translates the request to the LLM client's `format`, so use that format's field names, for example `reasoning_effort` on `openai_chat` or `reasoning` on `openai_responses`. Switchyard applies `extra_body` and `reasoning_effort` after the removal, so either can set a removed field again. |
152
153
|`reasoning_effort`| No | unset | Reasoning effort forced on every request to this target, replacing the value the caller sent (`reasoning.effort` on `openai_responses`, `reasoning_effort` on `openai_chat`). Rejected on `anthropic_messages` clients. Use it to run one target at a different effort than the client asked for, for example a strong tier at `max` behind a client that sends `high`. Targets with different effort settings need distinct model IDs when used within one route. Separate routes may use the same model ID with separate `llm_clients` entries (same endpoint, different name). |
153
154
154
155
Within one route, callable targets with the same model ID must use the same `llm_client`.
0 commit comments