[mod] ai_summary plugin: switch to the OpenAI chat completions API
Talk to the LLM server via GET /v1/models and POST /v1/chat/completions (SSE) instead of Ollama's native API. Any OpenAI compatible server now works (Ollama, vLLM, llama.cpp, LM Studio, Hugging Face TGI, ...); Ollama serves this API natively, existing setups keep working unchanged. The Ollama specific keep_alive option is dropped, the ai_summary.grounding setting is added as instance wide default of the grounding preference. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
This commit is contained in:
@@ -5,8 +5,10 @@
|
||||
===============
|
||||
|
||||
Default configuration of the :ref:`AI summary plugin <ai_summary plugin>`.
|
||||
Users configure the Ollama server URL and the model in the *AI Summary* tab of
|
||||
their preferences; the values below only act as instance wide defaults.
|
||||
Users configure the LLM server URL (any server implementing the OpenAI chat
|
||||
completions API: Ollama, vLLM, llama.cpp, LM Studio, Hugging Face TGI, ...)
|
||||
and the model in the *AI Summary* tab of their preferences; the values below
|
||||
only act as instance wide defaults.
|
||||
|
||||
.. code:: yaml
|
||||
|
||||
|
||||
Reference in New Issue
Block a user