dhi.io/vllm-openai
High-throughput LLM serving with an OpenAI-compatible API
pushed 2 days ago
linux/amd64, linux/arm64
4.76 GB
for linux/amd64
786
for linux/amd64
Active support
-
Build Level 3
Tamper-evident proof of how and from what sources this image was built. Build Level 3 is the highest the SLSA build track defines.
SPDX SBOM
A software bill of materials in both CycloneDX and SPDX formats so it drops straight into your existing tooling. Audit exactly what you're shipping.
-
A Vulnerability Exploitability exchange document, Docker's assessment of which CVEs actually affect this image and which don't apply, so you can focus on real risk instead of chasing false positives.
Available
A link to everything used to build the image, package source code, Git repos, and build files, so you can audit or reproduce the build and stay compliant with open source licenses.
Every known vulnerability in this image, shown in full rather than hidden. The VEX data flags which ones actually apply, so you can tell real exposure from noise before you ship.
-
Exactly what changed in this build, down to the package bumps and fixes behind the version you're pulling.
Verifies no keys, tokens, or credentials were accidentally baked into the image.
The image layers were scanned for known malware signatures before publishing.
vLLM is a high-throughput, memory-efficient inference and serving engine for large language models. The vLLM OpenAI server exposes OpenAI-compatible HTTP APIs for serving supported models on NVIDIA GPUs.
This image packages the vLLM OpenAI server with CUDA-enabled PyTorch and defaults to the upstream vllm serve
entrypoint.
Docker Hardened Images are built to meet the highest security and compliance standards. They provide a trusted foundation for containerized workloads by incorporating security best practices from the start.
These images are published with near-zero known CVEs, include signed provenance, and come with a complete Software Bill of Materials (SBOM) and VEX metadata. They're designed to secure your software supply chain while fitting seamlessly into existing Docker workflows.
vLLM® is a registered trademark of LF Projects, LLC. All rights in the mark are reserved to LF Projects, LLC. Any use by Docker is for referential purposes only and does not indicate sponsorship, endorsement, or affiliation.
NVIDIA® and CUDA® are trademarks and/or registered trademarks of NVIDIA Corporation in the U.S. and other countries.
This listing is prepared by Docker. All third-party product names, logos, and trademarks are the property of their respective owners and are used solely for identification. Docker claims no interest in those marks, and no affiliation, sponsorship, or endorsement is implied.