-
Notifications
You must be signed in to change notification settings - Fork 130
Pull requests: vllm-project/vllm-project.github.io
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
[Blog] Fix redundant 'vllm serve' in DGX Spark Docker invocation
#305
opened Aug 11, 2026 by
xinrub-droid
Loading…
[Blog] Keeping vLLM Fast Under CPU Pressure: An sched_ext Scheduler for GPU Inference
#300
opened Aug 7, 2026 by
ianchen0119
Loading…
Blog: Distributed Layerwise Offload — Running 185GB Models on 64GB NPU/GPU
#295
opened Jul 31, 2026 by
evanchueng
Contributor
Loading…
Blog: Add Office Hours #52 recap — Semantic Router and DeepLearning.AI course
#256
opened Jun 25, 2026 by
soyr-redhat
Loading…
Blog: Add Office Hours #51 recap — Speculators v0.5 and sparse MLA
#247
opened Jun 16, 2026 by
soyr-redhat
Loading…
Add Office Hours #50 recap: Running vLLM on Intel Xeon CPUs
#246
opened Jun 16, 2026 by
soyr-redhat
Loading…
Add blog post: Inside the vLLM-Omni Architecture: Serving Qwen3-Omni
#236
opened Jun 8, 2026 by
itigges22
Loading…
5 tasks done
ProTip!
Mix and match filters to narrow down what you’re looking for.