by yangyaofei · v0.2.3
vllm provider for extra_body support https://docs.vllm.ai/en/latest/serving/openai_compatible_server.html#id5
This community listing does not yet include every recommended support, privacy, pricing, and permission disclosure. Review the available package permissions before installing.
Available inside your emploidai workspace after installation.
Available inside your emploidai workspace after installation.
Dify custom model provider for vLLM's OpenAI-Compatible Server, supporting extra parameters and thinking mode features.
Based on the official Dify OpenAI-API-compatible plugin, extended for vLLM OpenAI-Compatible Server.
<think/>/</think/> → <think>/</think>, aligned 1:1 with dify-official-plugins openai_api_compatibleextra_body parameter delivery — JSON contents now merged into top-level request body instead of nested under "extra_body" keyenable_thinking toggle with compatibility_mode (strict/extended)reasoning_effort (none/low/medium/high), natively supported by vLLMchat_template_kwargs, thinking, enable_thinking at top levelresponse_format, json_schema, reasoning_format<think>...</think> when thinking is disabledreasoning (vLLM >= 0.17.1), fallback to reasoning_contentBreaking Change: Removed all legacy parameters, use
extra_bodyfor all extra parameter needs.
Same as OpenAI-API-compatible, select "Vllm" provider:
Pass extra parameters via the extra_body JSON text field:
Example:
{"chat_template_kwargs": {"enable_thinking": true}}