← Developers
Open on GitHub 


#1125 2
Michael Goin
@mgoin · USA
17.9k
Weighted score
2.09k
Contributions
16
Repos
Top repos
- vllm-project/vllm
- weicj/vLLM-2080Ti-Definitive
- 1CatAI/1Cat-vLLM
- neuralmagic/deepsparse
- vllm-project/llm-compressor
- neuralmagic/sparseml
- EleutherAI/lm-evaluation-harness
- vllm-project/speculators
- vllm-project/vllm-metal
- rom1504/clip-retrieval
Ranked AI repos3
23531
vllm-project/vllm
+51
A high-throughput and memory-efficient inference and serving engine for LLMs
93.2k· Python· Infrastructure
167671635
neuralmagic/deepsparse
+-1
Sparsity-aware deep learning inference runtime for CPUs
3.15k· Python· Infrastructure
1245222
weicj/vLLM-2080Ti-Definitive
+7
The definitive vLLM runtime for dual RTX 2080 Ti 22GB + NVLink, delivering Qwen 27B local inference with maximum 200+ tok/s single-request decode with support of FP8 weight ( Join Discord :https://discord.gg/VFqVVySdMS )
1.09k· Python· Infrastructure