1
[Feature Request] Allow configuring multiple LLM endpoints with priority and automatic failover.
Source: paperless-ngx/paperless-ngx#14005 · opened by @fabiencharrasse
Description
This would allow users to use a more powerful Ollama server when available, while automatically falling back to another server if the preferred endpoint is offline or unreachable.
Example use case:
• Primary endpoint: a personal PC with a powerful GPU, only available when powered on.
• Fallback endpoint: a dedicated server with a less powerful GPU, always available.
This would allow Paperless to benefit from available hardware without requiring an external proxy or manual configuration changes.
Other
_No response_
This would allow users to use a more powerful Ollama server when available, while automatically falling back to another server if the preferred endpoint is offline or unreachable.
Example use case:
• Primary endpoint: a personal PC with a powerful GPU, only available when powered on.
• Fallback endpoint: a dedicated server with a less powerful GPU, always available.
This would allow Paperless to benefit from available hardware without requiring an external proxy or manual configuration changes.
Other
_No response_
No pledges yet. Be the first to back this.
Comments
Similar requests
[Feature Request] Different keys for LLM vs. LLM embeddings
6 votes · 0 comments
[Feature Request] Automatically suggest non-LLM suggestions when editing documents, like in versions before 3.0.0
15 votes · 0 comments
[Feature Request] Allow overriding the embeddings encoding_format for OpenAI-compatible LLM providers
4 votes · 0 comments
[Feature Request] Logging for AI/LLM usage
4 votes · 0 comments
[Feature Request] AI Suggestions Cache
6 votes · 0 comments
No comments yet.