github.com/vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
https://github.com/vllm-project/vllm
nixpkgs: python311Packages.vllm 0.3.3
A high-throughput and memory-efficient inference and serving engine for LLMs1 version - Latest release: 6 months ago
helm: infracloud-charts/vllm 0.2.2
A Helm chart for deploying models with vLLM4 versions - Latest release: over 1 year ago - 0 downloads total
nixpkgs: python313Packages.vllm 0.15.1
High-throughput and memory-efficient inference and serving engine for LLMs2 versions - Latest release: 5 months ago
nixpkgs: python312Packages.vllm 0.3.3
A high-throughput and memory-efficient inference and serving engine for LLMs1 version - Latest release: 6 months ago
pypi: byzerllm 0.1.182
ByzerLLM: Byzer LLM178 versions - Latest release: over 1 year ago - 1 dependent package - 2 dependent repositories - 8.85 thousand downloads last month
pypi: personal-test-vllm-tpu 0.10.1.1
A high-throughput and memory-efficient inference and serving engine for LLMs1 version - Latest release: 11 months ago - 0 downloads last month
pypi: tmp-test-vllm-tpu 0.0.1
A high-throughput and memory-efficient inference and serving engine for LLMs1 version - Latest release: 11 months ago - 168 downloads last month
pypi: vllm-test-tpu 0.9.1
A high-throughput and memory-efficient inference and serving engine for LLMs2 versions - Latest release: about 1 year ago - 21 downloads last month
pypi: vllm-fixed 1.0.0
A high-throughput and memory-efficient inference and serving engine for LLMs2 versions - Latest release: 10 months ago - 15 downloads last month
pypi: llm-engines 0.0.17
A unified inference engine for large language models (LLMs) including open-source models (VLLM, S...17 versions - Latest release: almost 2 years ago - 506 downloads last month
pypi: vllm-cpu-nightly 0.26.1.dev202608050841
A high-throughput and memory-efficient inference and serving engine for LLMs62 versions - Latest release: 9 days ago - 12.6 thousand downloads last month
pypi: vllm-hust 0.17.2.post1
A high-throughput and memory-efficient inference and serving engine for LLMs4 versions - Latest release: 4 months ago - 38 downloads last month
pypi: vllm-musa 0.1.1
vLLM platform plugin for Moore Threads MUSA GPUs3 versions - Latest release: 8 months ago - 68 downloads last month
pypi: vllm-usf 0.0.2
vLLM-USF: A high-throughput and memory-efficient inference engine for LLMs (USF Custom Build)2 versions - Latest release: 10 months ago - 27 downloads last month
conda: vllm 0.21.0
vLLM is a fast and easy-to-use library for LLM inference and serving.2 versions - Latest release: 23 days ago - 110 downloads total
pypi: vllm-tpu 0.26.0
A high-throughput and memory-efficient inference and serving engine for LLMs27 versions - Latest release: 14 days ago - 48.8 thousand downloads last month
pypi: mindie-turbo 2.0rc1
MindIE Turbo: An LLM inference acceleration framework featuring extensive plugin collections opti...1 version - Latest release: over 1 year ago - 187 downloads last month
pypi: vllm-emissary 0.1.0
A high-throughput and memory-efficient inference and serving engine for LLMs2 versions - Latest release: over 1 year ago - 28 downloads last month
pypi: ai-dynamo-vllm 0.8.4
A high-throughput and memory-efficient inference and serving engine for LLMs7 versions - Latest release: over 1 year ago - 86 downloads last month
pypi: wxy-test 0.19.0
A high-throughput and memory-efficient inference and serving engine for LLMs3 versions - Latest release: 4 months ago - 10 downloads last month
pypi: vllm-npu 0.4.2
A high-throughput and memory-efficient inference and serving engine for LLMs3 versions - Latest release: over 1 year ago - 40 downloads last month
pypi: vllm-rocm 0.6.3
A high-throughput and memory-efficient inference and serving engine for LLMs with AMD GPU support1 version - Latest release: almost 2 years ago - 68 downloads last month
pypi: llm_math 0.2.0
A tool designed to evaluate the performance of large language models on mathematical tasks.5 versions - Latest release: almost 2 years ago - 114 downloads last month
pypi: moe-kernels 0.8.2
MoE kernels15 versions - Latest release: over 1 year ago - 181 downloads last month
Top 3.4% on pypi.org
93 versions - Latest release: 20 days ago - 46 dependent packages - 5 dependent repositories - 5.5 million downloads last month
pypi: vllm 0.26.0
A high-throughput and memory-efficient inference and serving engine for LLMs93 versions - Latest release: 20 days ago - 46 dependent packages - 5 dependent repositories - 5.5 million downloads last month
pypi: marlin-kernels 0.3.7
Marlin quantization kernels11 versions - Latest release: over 1 year ago - 159 downloads last month
pypi: vllm-acc 0.4.1
A high-throughput and memory-efficient inference and serving engine for LLMs8 versions - Latest release: over 2 years ago - 60 downloads last month
pypi: vllm-online 0.4.2
A high-throughput and memory-efficient inference and serving engine for LLMs2 versions - Latest release: over 2 years ago - 26 downloads last month
pypi: tilearn-infer 0.3.3
A high-throughput and memory-efficient inference and serving engine for LLMs3 versions - Latest release: over 2 years ago - 19 downloads last month
pypi: tilearn-test01 0.1
A high-throughput and memory-efficient inference and serving engine for LLMs1 version - Latest release: over 2 years ago - 15 downloads last month
pypi: llm_atc 0.1.7
Tools for fine tuning and serving LLMs6 versions - Latest release: over 2 years ago - 193 downloads last month
pypi: vllm-xft 0.5.5.4
A high-throughput and memory-efficient inference and serving engine for LLMs12 versions - Latest release: over 1 year ago - 31 downloads last month
pypi: superlaser 0.0.6
An MLOps library for LLM deployment w/ the vLLM engine on RunPod's infra.6 versions - Latest release: over 2 years ago - 180 downloads last month
pypi: llm-swarm 0.1.1
A high-throughput and memory-efficient inference and serving engine for LLMs2 versions - Latest release: over 2 years ago - 72 downloads last month
pypi: nextai-vllm 0.0.7
A high-throughput and memory-efficient inference and serving engine for LLMs6 versions - Latest release: over 2 years ago - 35 downloads last month
pypi: vllm-consul 0.2.1
A high-throughput and memory-efficient inference and serving engine for LLMs5 versions - Latest release: almost 3 years ago - 39 downloads last month
Top 9.6% on proxy.golang.org
81 versions - Latest release: 20 days ago
go: github.com/vllm-project/vllm v0.26.0
A high-throughput and memory-efficient inference and serving engine for LLMs81 versions - Latest release: 20 days ago
nixpkgs: python314Packages.vllm 0.15.1
High-throughput and memory-efficient inference and serving engine for LLMs2 versions - Latest release: 5 months ago
nixpkgs: vllm 0.15.1
High-throughput and memory-efficient inference and serving engine for LLMs2 versions - Latest release: 5 months ago