v0.9.1 #138
nvluxiaoz
announced in
Announcements
v0.9.1
#138
Replies: 0 comments
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Uh oh!
There was an error while loading. Please reload this page.
TensorRT Edge-LLM 0.9.1 Release 07/23/2026
We are excited to announce the release 0.9.1 of TensorRT Edge-LLM!
TensorRT Edge-LLM 0.9.1 expands Gemma 4 multimodal support, DFlash, paged KV cache, and OpenAI-compatible generation controls while improving runtime performance and stability.
Key Features
Other Important Features
Runtime and Performance
Export and Quantization
Server and API
Examples
NVIDIA Contributors
@nvluxiaoz @nvamberl @xiangg-nv @willg-nv @mahu888 @nv-samcheng @duofant @Caohanwen0 @JCalafato @zhazhang-nv @zhaoyuanh-nvidia @zhijial-nvidia @ever-wong @Jasper-NV @wanghr323 @jhalabi-nv @levichen-nvidia @sunghyunp-nvidia @nvyocox @fans-nv @ruocheng-nv @yuanyao-nv @duanyaqi
This discussion was created from the release v0.9.1.
All reactions