vllm 0.0.2
Chart · Machine learning & AI · standard · v0.0.2
vLLM, the high-throughput LLM inference and serving engine with an OpenAI-compatible API. CPU build — installs vLLM's prebuilt CPU wheel (torch+cpu, no source compile, no CUDA) into a venv on a hardened Wolfi python-3.12 base, 0-CVE across ~140 Python packages. The user supplies the model; the chart runs it as a StatefulSet with a persistent model cache.
Version
The latest line lives at the base page; older lines have their own page so you can pin and verify exactly that version.
Released version
This is the 0.0.2 release of the vllm chart, published 2026-07-09.For the live security report and the currently deployed image digest, see the latest release.
Security report (Trivy)
Security report (Trivy) · image vllm 0.25.1
Install the chart
Deploy to Kubernetes with hardened defaults. The chart pins its image by signed digest, so you never track it yourself.
Install (pinned to 0.0.2)
helm install my-vllm oci://ghcr.io/quenchworks/charts/vllm --version 0.0.2- Chart version
- 0.0.2
- Chart license
- Apache-2.0
- App license
- Apache-2.0
- Service port
- 8000
- Signed
- cosign (keyless)
- Released
- 2026-07-09
Verify the chart
cosign verify ghcr.io/quenchworks/charts/vllm:0.0.2 \
--certificate-identity-regexp 'https://github.com/quenchworks/.+' \
--certificate-oidc-issuer https://token.actions.githubusercontent.comTransparency
The chart publishes its attestations on GitHub and the image it deploys carries its own on the same digest, publicly verifiable with the commands above. Both log to the Sigstore transparency log (Rekor), which cosign verify checks for you.
Upstream project: https://docs.vllm.ai