vLLM is an inference and serving engine for large language models. Prior to 0.27.0, an integer overflow in blockIdx.x * 2 * d in activation_kernels.cu can cause act_and_mul_kernel to consume another batched user's input, allowing a request processed in the same inference batch to receive a partial or complete copy of another user's inference result. This issue is fixed in version 0.27.0.
References
Configurations
No configuration.
History
No history.
Information
Published : 2026-08-13 15:20
Updated : 2026-09-09 20:58
NVD link : CVE-2026-73558
Mitre link : CVE-2026-73558
CVE.ORG link : CVE-2026-73558
JSON object : View
Products Affected
No product.
CWE
CWE-190
Integer Overflow or Wraparound
