App
github-com-vllm-project-vllm
x
Search
Install Pinokio
Log in
Register
Log in
Register
Light mode
vllm
https://github.com/vllm-project/vllm
updated 2/12/2026, 8:22:06 AM
indexed 2/12/2026, 10:33:28 AM
A high-throughput and memory-efficient inference and serving engine for LLMs
Follow
0
Loading community details…
Community
Search this app
Sort by
Recent activity
Newest
Top: Past day
Top: Past week
Top: Past month
Top: Past year
Top: All
Post about vllm...
Post
Loading...
Pinokio Apps Using This Repo
No Pinokio apps using this repo yet.
Community tags
None yet.
Check-ins
(0)
View all
Platforms (0)
No reports yet.
Arch (0)
No reports yet.
GPU (0)
No reports yet.
RAM (0)
No reports yet.
VRAM (0)
No reports yet.
Recent commits
Vllm CPU benchmark suite improvement (#34128)
Louie Tsai
7 months ago
55a1a95
[Bug Fix] Fix `naive_block_assignment` always defaulting to False due to arg misalignment (#33848)
Runkai Tao
7 months ago
e1d97c3
[Bugfix] Fix MTP accuracy for GLM-5 (#34385)
Michael Goin
7 months ago
ec12d39
Fix Mistral config remap to accept compressed-tensors quantization #34028 (#34104)
baonudesifeizhai
7 months ago
f589761
[bugfix] refactor FunASR's _get_data_parser (#34397)
AllenDou
7 months ago
386bfe5
[BugFix] Fix DP chunking (#34379)
Lucas Wilkinson
7 months ago
136b0bf
[Bugfix] Fix Sparse24 Compressed Tensors models (#33446)
Kyle Sayers
7 months ago
e9cd691
Fix DeepSeek-OCR tensor validation for all size variants (#34085)
Yichuan Wang
7 months ago
80f2ba6
[Refactor] Pass Renderer to Input Processor (#34329)
Cyrus Leung
7 months ago
b96f731
[Refactor] Move validation to params definitions (#34362)
Cyrus Leung
7 months ago
ced2a92