OCR uses two threads by default #7685

Closed
opened 2026-02-05 13:13:48 +03:00 by OVERLORD · 1 comment
Owner

Originally created by @wnhre2ur8cxx8 on GitHub (Nov 1, 2025).

I have searched the existing issues, both open and closed, to make sure this is not a duplicate report.

  • Yes

The bug

I am running 2.2.1 and since 2.2.0 OCR uses 2 threads by default. If I configure it to run 2 jobs in parallell it will add another one, so it runs on three threads instead of the expected two. I am using the server model.

Image

The OS that Immich Server is running on

Debian 13

Version of Immich Server

2.2.1

Version of Immich Mobile App

2.2.1

Platform with the issue

  • Server
  • Web
  • Mobile

Device make and model

No response

Your docker-compose.yml content

#
# WARNING: Make sure to use the docker-compose.yml of the current release:
#
# https://github.com/immich-app/immich/releases/latest/download/docker-compose.yml
#
# The compose file on main may not be compatible with the latest release.
#

name: immich

services:
  immich-server:
    container_name: immich_server
    image: ghcr.io/immich-app/immich-server:${IMMICH_VERSION:-release}
    # extends:
    #   file: hwaccel.transcoding.yml
    #   service: cpu # set to one of [nvenc, quicksync, rkmpp, vaapi, vaapi-wsl] for accelerated transcoding
    volumes:
      # Do not edit the next line. If you want to change the media storage location on your system, edit the value of UPLOAD_LOCATION in the .env file
      - ${UPLOAD_LOCATION}:/data
      - /etc/localtime:/etc/localtime:ro
    env_file:
      - .env
    ports:
      - '127.0.0.1:2283:2283'
    depends_on:
      - redis
      - database
    restart: always
    healthcheck:
      disable: false
    logging:
      driver: "journald"
      options:
        tag: "immich-server"

  immich-machine-learning:
    container_name: immich_machine_learning
    # For hardware acceleration, add one of -[armnn, cuda, openvino] to the image tag.
    # Example tag: ${IMMICH_VERSION:-release}-cuda
    image: ghcr.io/immich-app/immich-machine-learning:${IMMICH_VERSION:-release}
    # extends: # uncomment this section for hardware acceleration - see https://immich.app/docs/features/ml-hardware-acceleration
    #   file: hwaccel.ml.yml
    #   service: cpu # set to one of [armnn, cuda, openvino, openvino-wsl] for accelerated inference - use the `-wsl` version for WSL2 where applicable
    volumes:
      - model-cache:/cache
    env_file:
      - .env
    restart: always
    healthcheck:
      disable: false

  redis:
    container_name: immich_redis
    image: docker.io/valkey/valkey:8@sha256:81db6d39e1bba3b3ff32bd3a1b19a6d69690f94a3954ec131277b9a26b95b3aa
    healthcheck:
      test: redis-cli ping || exit 1
    restart: always

  database:
    container_name: immich_postgres
    image: ghcr.io/immich-app/postgres:14-vectorchord0.4.3-pgvectors0.2.0@sha256:bcf63357191b76a916ae5eb93464d65c07511da41e3bf7a8416db519b40b1c23
    environment:
      POSTGRES_PASSWORD: ${DB_PASSWORD}
      POSTGRES_USER: ${DB_USERNAME}
      POSTGRES_DB: ${DB_DATABASE_NAME}
      POSTGRES_INITDB_ARGS: '--data-checksums'
      # Uncomment the DB_STORAGE_TYPE: 'HDD' var if your database isn't stored on SSDs
      DB_STORAGE_TYPE: 'HDD'
    volumes:
      # Do not edit the next line. If you want to change the database storage location on your system, edit the value of DB_DATA_LOCATION in the .env file
      - ${DB_DATA_LOCATION}:/var/lib/postgresql/data
    shm_size: 128mb
    restart: always

volumes:
  model-cache:

networks:
  default:
    driver: bridge
    ipam:
      config:
        - subnet: 172.100.0.0/16

Your .env content

# You can find documentation for all the supported env variables at https://immich.app/docs/install/environment-variables

# The location where your uploaded files are stored
UPLOAD_LOCATION=/data/immich-library
# The location where your database files are stored
DB_DATA_LOCATION=/root/immich-data/postgres-data

# To set a timezone, uncomment the next line and change Etc/UTC to a TZ identifier from this list: https://en.wikipedia.org/wiki/List_of_tz_database_time_zones#List
TZ=Europe/Berlin

# The Immich version to use. You can pin this to a specific version like "v1.71.0"
IMMICH_VERSION=v2

# Connection secret for postgres. You should change it to a random password
# Please use only the characters `A-Za-z0-9`, without special characters or spaces
DB_PASSWORD=stripped

# The values below this line do not need to be changed
###################################################################################
DB_USERNAME=postgres
DB_DATABASE_NAME=immich

MACHINE_LEARNING_PRELOAD__CLIP__TEXTUAL=ViT-SO400M-16-SigLIP2-384__webli

Reproduction steps

  1. Use server OCR model
  2. set OCR parallell jobs to 1
  3. start OCR queue
  4. check cpu load
  5. see 2 active ML threads
  6. stop OCR queue
  7. see no active ML threads

Relevant log output

[10/31/25 23:07:40] INFO     Starting gunicorn 23.0.0                           
[10/31/25 23:07:40] INFO     Listening at: http://[::]:3003 (8)                 
[10/31/25 23:07:40] INFO     Using worker: immich_ml.config.CustomUvicornWorker 
[10/31/25 23:07:40] INFO     Booting worker with pid: 9                         
[10/31/25 23:07:47] INFO     Started server process [9]                         
[10/31/25 23:07:47] INFO     Waiting for application startup.                   
[10/31/25 23:07:47] INFO     Created in-memory cache with unloading after 300s  
                             of inactivity.                                     
[10/31/25 23:07:47] INFO     Initialized request thread pool with 4 threads.    
[10/31/25 23:07:47] INFO     Preloading models:                                 
                             clip:textual='ViT-SO400M-16-SigLIP2-384__webli'    
                             visual=None facial_recognition:recognition=None    
                             detection=None                                     
[10/31/25 23:07:47] INFO     Loading textual model                              
                             'ViT-SO400M-16-SigLIP2-384__webli' to memory       
[10/31/25 23:07:47] INFO     Setting execution providers to                     
                             ['CPUExecutionProvider'], in descending order of   
                             preference                                         
[10/31/25 23:07:52] INFO     Application startup complete.                      
[10/31/25 23:07:52] INFO     Loading detection model 'PP-OCRv5_server' to memory
[10/31/25 23:07:52] INFO     Setting execution providers to                     
                             ['CPUExecutionProvider'], in descending order of   
                             preference                                         
[10/31/25 23:07:53] INFO     Using engine_name: onnxruntime                     
[10/31/25 23:08:17] INFO     Loading recognition model 'PP-OCRv5_server' to     
                             memory                                             
[10/31/25 23:08:17] INFO     Setting execution providers to                     
                             ['CPUExecutionProvider'], in descending order of   
                             preference                                         
[10/31/25 23:08:18] INFO     Using engine_name: onnxruntime

Additional information

No response

Originally created by @wnhre2ur8cxx8 on GitHub (Nov 1, 2025). ### I have searched the existing issues, both open and closed, to make sure this is not a duplicate report. - [x] Yes ### The bug I am running 2.2.1 and since 2.2.0 OCR uses 2 threads by default. If I configure it to run 2 jobs in parallell it will add another one, so it runs on three threads instead of the expected two. I am using the server model. <img width="2256" height="376" alt="Image" src="https://github.com/user-attachments/assets/e709af75-a778-44a3-b935-e0c1d6fd0ac1" /> ### The OS that Immich Server is running on Debian 13 ### Version of Immich Server 2.2.1 ### Version of Immich Mobile App 2.2.1 ### Platform with the issue - [x] Server - [ ] Web - [ ] Mobile ### Device make and model _No response_ ### Your docker-compose.yml content ```YAML # # WARNING: Make sure to use the docker-compose.yml of the current release: # # https://github.com/immich-app/immich/releases/latest/download/docker-compose.yml # # The compose file on main may not be compatible with the latest release. # name: immich services: immich-server: container_name: immich_server image: ghcr.io/immich-app/immich-server:${IMMICH_VERSION:-release} # extends: # file: hwaccel.transcoding.yml # service: cpu # set to one of [nvenc, quicksync, rkmpp, vaapi, vaapi-wsl] for accelerated transcoding volumes: # Do not edit the next line. If you want to change the media storage location on your system, edit the value of UPLOAD_LOCATION in the .env file - ${UPLOAD_LOCATION}:/data - /etc/localtime:/etc/localtime:ro env_file: - .env ports: - '127.0.0.1:2283:2283' depends_on: - redis - database restart: always healthcheck: disable: false logging: driver: "journald" options: tag: "immich-server" immich-machine-learning: container_name: immich_machine_learning # For hardware acceleration, add one of -[armnn, cuda, openvino] to the image tag. # Example tag: ${IMMICH_VERSION:-release}-cuda image: ghcr.io/immich-app/immich-machine-learning:${IMMICH_VERSION:-release} # extends: # uncomment this section for hardware acceleration - see https://immich.app/docs/features/ml-hardware-acceleration # file: hwaccel.ml.yml # service: cpu # set to one of [armnn, cuda, openvino, openvino-wsl] for accelerated inference - use the `-wsl` version for WSL2 where applicable volumes: - model-cache:/cache env_file: - .env restart: always healthcheck: disable: false redis: container_name: immich_redis image: docker.io/valkey/valkey:8@sha256:81db6d39e1bba3b3ff32bd3a1b19a6d69690f94a3954ec131277b9a26b95b3aa healthcheck: test: redis-cli ping || exit 1 restart: always database: container_name: immich_postgres image: ghcr.io/immich-app/postgres:14-vectorchord0.4.3-pgvectors0.2.0@sha256:bcf63357191b76a916ae5eb93464d65c07511da41e3bf7a8416db519b40b1c23 environment: POSTGRES_PASSWORD: ${DB_PASSWORD} POSTGRES_USER: ${DB_USERNAME} POSTGRES_DB: ${DB_DATABASE_NAME} POSTGRES_INITDB_ARGS: '--data-checksums' # Uncomment the DB_STORAGE_TYPE: 'HDD' var if your database isn't stored on SSDs DB_STORAGE_TYPE: 'HDD' volumes: # Do not edit the next line. If you want to change the database storage location on your system, edit the value of DB_DATA_LOCATION in the .env file - ${DB_DATA_LOCATION}:/var/lib/postgresql/data shm_size: 128mb restart: always volumes: model-cache: networks: default: driver: bridge ipam: config: - subnet: 172.100.0.0/16 ``` ### Your .env content ```Shell # You can find documentation for all the supported env variables at https://immich.app/docs/install/environment-variables # The location where your uploaded files are stored UPLOAD_LOCATION=/data/immich-library # The location where your database files are stored DB_DATA_LOCATION=/root/immich-data/postgres-data # To set a timezone, uncomment the next line and change Etc/UTC to a TZ identifier from this list: https://en.wikipedia.org/wiki/List_of_tz_database_time_zones#List TZ=Europe/Berlin # The Immich version to use. You can pin this to a specific version like "v1.71.0" IMMICH_VERSION=v2 # Connection secret for postgres. You should change it to a random password # Please use only the characters `A-Za-z0-9`, without special characters or spaces DB_PASSWORD=stripped # The values below this line do not need to be changed ################################################################################### DB_USERNAME=postgres DB_DATABASE_NAME=immich MACHINE_LEARNING_PRELOAD__CLIP__TEXTUAL=ViT-SO400M-16-SigLIP2-384__webli ``` ### Reproduction steps 1. Use server OCR model 2. set OCR parallell jobs to 1 3. start OCR queue 4. check cpu load 5. see 2 active ML threads 6. stop OCR queue 7. see no active ML threads ### Relevant log output ```shell [10/31/25 23:07:40] INFO Starting gunicorn 23.0.0 [10/31/25 23:07:40] INFO Listening at: http://[::]:3003 (8) [10/31/25 23:07:40] INFO Using worker: immich_ml.config.CustomUvicornWorker [10/31/25 23:07:40] INFO Booting worker with pid: 9 [10/31/25 23:07:47] INFO Started server process [9] [10/31/25 23:07:47] INFO Waiting for application startup. [10/31/25 23:07:47] INFO Created in-memory cache with unloading after 300s of inactivity. [10/31/25 23:07:47] INFO Initialized request thread pool with 4 threads. [10/31/25 23:07:47] INFO Preloading models: clip:textual='ViT-SO400M-16-SigLIP2-384__webli' visual=None facial_recognition:recognition=None detection=None [10/31/25 23:07:47] INFO Loading textual model 'ViT-SO400M-16-SigLIP2-384__webli' to memory [10/31/25 23:07:47] INFO Setting execution providers to ['CPUExecutionProvider'], in descending order of preference [10/31/25 23:07:52] INFO Application startup complete. [10/31/25 23:07:52] INFO Loading detection model 'PP-OCRv5_server' to memory [10/31/25 23:07:52] INFO Setting execution providers to ['CPUExecutionProvider'], in descending order of preference [10/31/25 23:07:53] INFO Using engine_name: onnxruntime [10/31/25 23:08:17] INFO Loading recognition model 'PP-OCRv5_server' to memory [10/31/25 23:08:17] INFO Setting execution providers to ['CPUExecutionProvider'], in descending order of preference [10/31/25 23:08:18] INFO Using engine_name: onnxruntime ``` ### Additional information _No response_
Author
Owner

@mertalev commented on GitHub (Nov 1, 2025):

It's interesting that 1 active job uses 2 threads while 2 active jobs use 3. Jobs and threads aren't necessarily 1:1 though, so I don't think this is really a bug.

@mertalev commented on GitHub (Nov 1, 2025): It's interesting that 1 active job uses 2 threads while 2 active jobs use 3. Jobs and threads aren't necessarily 1:1 though, so I don't think this is really a bug.
Sign in to join this conversation.
1 Participants
Notifications
Due Date
No due date set.
Dependencies

No dependencies set.

Reference: immich-app/immich#7685