Skip to content

v0.2.1

Latest

Choose a tag to compare

@AnkurAlpha AnkurAlpha released this 05 Sep 16:44

Fix server lifecycle handling while inference providers are loading. FIMLoom now reuses responding endpoints, including llama.cpp HTTP 503 responses, without launching duplicate servers.