Fix server lifecycle handling while inference providers are loading. FIMLoom now reuses responding endpoints, including llama.cpp HTTP 503 responses, without launching duplicate servers.
Fix server lifecycle handling while inference providers are loading. FIMLoom now reuses responding endpoints, including llama.cpp HTTP 503 responses, without launching duplicate servers.