VulnWatch VulnWatch
← Back to dashboard
Medium nvd · CVE-2026-54235

CVE-2026-54235: vLLM is an inference and serving engine for large language models (LLMs). Prior to 0.23.1rc0, ll temperature validation

Published Jun 22, 2026 CVSS 6.9

vLLM is an inference and serving engine for large language models (LLMs). Prior to 0.23.1rc0, ll temperature validation gates use comparison operators (), which silently evaluate to False for NaN and for positive Infinity in Python's IEEE 754 float semantics. Both values pass every guard and propagate to GPU sampling kernels, where they produce undefined behavior or CUDA errors that can crash the inference worker. This vulnerability is fixed in 0.23.1rc0.

Affected AI Products

large language model vllm llm
Get the weekly digest. Every Monday: top AI security stories of the week. Free.