The service stays connected full time; the GPU is held only while translating. Loads at startup (backlog is likely after a downtime).
The service stays connected full time; the GPU is held only while translating. Loads at startup (backlog is likely after a downtime).