mirror of
https://github.com/tiennm99/litellm.git
synced 2026-08-02 06:22:48 +00:00
* docs: add OpenAI-compatible API limitations for Anthropic thinking Document the fundamental incompatibility between Anthropic extended thinking and OpenAI-compatible API clients. Explains: - Why thinking_blocks must be resent (stateless vs stateful APIs) - OpenAI vs Anthropic architecture differences - Solutions for client developers * Update docs * fix: auto-drop thinking param when thinking_blocks missing When modify_params=True, LiteLLM now automatically drops the 'thinking' param if the last assistant message with tool_calls is missing thinking_blocks. This prevents the Anthropic error: "Expected thinking or redacted_thinking, but found tool_use" This workaround addresses the OpenAI-Anthropic API incompatibility where OpenAI-compatible clients don't preserve thinking_blocks.
Website
This website is built using Docusaurus 2, a modern static website generator.
Installation
$ yarn
Local Development
$ yarn start
This command starts a local development server and opens up a browser window. Most changes are reflected live without having to restart the server.
Build
$ yarn build
This command generates static content into the build directory and can be served using any static contents hosting service.
Deployment
Using SSH:
$ USE_SSH=true yarn deploy
Not using SSH:
$ GIT_USER=<Your GitHub username> yarn deploy
If you are using GitHub pages for hosting, this command is a convenient way to build the website and push to the gh-pages branch.