1
Improve machine learning resilience in low VRAM scenarios (CUDA)
Source: immich-app/immich#11981 · opened by @flobernd
The bug Hi there, I'm currently evaluating Immich and really liking it so far. My Immich instance is running via Docker on a Debian 12 VM that is hosted on ESXi. The VM has a vGPU profile with 4 GiB VRAM assigned. During a stress-test of the hardware accelleration (both, transcoding and machine learning), I noticed that the machine learning Python module does not seem to be very resilient against low VRAM scenarios. After uploading some initial photos and videos, I run a "stress-test" by starting re-running all relevant processing tasks simultaneously (face detection, smart search, transcoding). At the same time, I ran a smart search query from the main dashboard. The following observations were made: • Especially the smart search query allocates a lot of VRAM • When the smart search query request fails due to low memory: 1. A corresponding exception (failed to allocate memory) is logged in the container 2. The smart search container…
No pledges yet. Be the first to back this.
Comments
Similar requests
[Feature] Upgrade to ONNX Runtime 1.27.0 for official CUDA 13 support
1 vote · 0 comments
[Feature] Support for arm64 for CUDA (immich-machine-learning)
41 votes · 0 comments
[Feature] Separate Image Preview for Machine Learning
1 vote · 0 comments
[Feature] DirectML execution provider for Machine Learning (AMD GPU Support on Windows/WSL2)
8 votes · 0 comments
[Feature] Automatically pause and resume Remote Machine Learning for on-and-off beefy machine
18 votes · 0 comments
No comments yet.