Skip to content

swiss-ai/vllm

Error
Looks like something went wrong!

About

A high-throughput and memory-efficient inference and serving engine for LLMs

Resources

License

Code of conduct

Security policy

Stars

Watchers

Forks

Packages

No packages published

Languages

  • Python 84.8%
  • Cuda 10.2%
  • C++ 3.5%
  • C 0.6%
  • Shell 0.5%
  • CMake 0.3%
  • Dockerfile 0.1%