vllm-stack-0.0.4
What's Changed
- [Doc] Update README.md by @Shaoting-Feng in #73
- feat: OpenAI batch API part 1 by @gaocegege in #52
- [Add] fix for router files api and example to post query to the api by @ApostaC in #76
- feat: Add basic issue templates by @gaocegege in #79
- chore: Add test cases for file storage by @gaocegege in #82
- feat: Wrap router to a singleton by @gaocegege in #83
- adding step-by-step tutorial links in readme by @junchenj in #84
- [Doc] Fix README section by @Shaoting-Feng in #85
- chore: Refine README, adjust image size by @gaocegege in #88
- [Doc] Add PR template by @Shaoting-Feng in #93
- feat(router): generate req id with uuid. by @Electronic-Waste in #89
- Feat: Add support for disabling router by @0xThresh in #96
- Update yaml file for the tutorials by @junchenj in #98
- [CI/Build] : add GitHub Actions workflows for router (#74) by @Sozhan308 in #94
- [CI/Build] Add helm update to helm func test pipeline by @Shaoting-Feng in #99
- [CI/Build] Avoid using helm repo by @Shaoting-Feng in #100
- Enable multi-GPU inference in vLLM with tensor parallelism by @YuhanLiu11 in #105
New Contributors
- @junchenj made their first contribution in #84
- @Electronic-Waste made their first contribution in #89
- @Sozhan308 made their first contribution in #94
- @YuhanLiu11 made their first contribution in #105
Full Changelog: vllm-stack-0.0.3...vllm-stack-0.0.4