← Developers
Open on GitHub 


#850 3
Woosuk Kwon
@WoosukKwon · USA
21.8k
Weighted score
2.54k
Contributions
5
Repos
Top repos
- vllm-project/vllm
- weicj/vLLM-2080Ti-Definitive
- 1CatAI/1Cat-vLLM
- skypilot-org/skypilot
- alpa-projects/alpa
Ranked AI repos3
31244
vllm-project/vllm
+52
A high-throughput and memory-efficient inference and serving engine for LLMs
93.4k· Python· Infrastructure
2182732
1CatAI/1Cat-vLLM
+4
V100 / SM70-focused vLLM engineering fork for modern LLM inference.
1.26k· Python· Infrastructure
2200594
weicj/vLLM-2080Ti-Definitive
+4
The definitive vLLM runtime for dual RTX 2080 Ti 22GB + NVLink, delivering Qwen 27B local inference with maximum 200+ tok/s single-request decode with support of FP8 weight ( Join Discord :https://discord.gg/VFqVVySdMS )
1.11k· Python· Infrastructure