<?xml version='1.0' encoding='UTF-8'?>
<?xml-stylesheet href="/static/style.xsl" type="text/xsl"?>
<feed xmlns="http://www.w3.org/2005/Atom" xml:lang="en">
  <id>https://cve.radiocsirt.org/rss/recent/all/10</id>
  <title>Most recent entries from all</title>
  <updated>2026-10-03T03:32:24.394802+00:00</updated>
  <author>
    <name>Vulnerability-Lookup</name>
    <email>csirt@opendfir.org</email>
  </author>
  <link href="https://cve.radiocsirt.org" rel="alternate"/>
  <generator uri="https://lkiesow.github.io/python-feedgen" version="1.0.0">python-feedgen</generator>
  <subtitle>Contains only the most 10 recent entries.</subtitle>
  <entry>
    <id>https://cve.radiocsirt.org/vuln/cve-2026-54235</id>
    <title>CVE-2026-54235 — vLLM: temperature=NaN and temperature=Infinity bypass validation and propagate to GPU kernels</title>
    <updated>2026-10-03T03:32:24.425725+00:00</updated>
    <content type="xhtml">
      <div xmlns="http://www.w3.org/1999/xhtml"><p><strong>Affected:</strong> vllm-project vllm</p>
<p>vLLM is an inference and serving engine for large language models (LLMs). Prior to 0.23.1rc0, ll temperature validation gates use comparison operators (&lt;, &gt;), which silently evaluate to False for NaN and for positive Infinity in Python's IEEE 754 float semantics. Both values pass every guard and propagate to GPU sampling kernels, where they produce undefined behavior or CUDA errors that can crash the inference worker. This vulnerability is fixed in 0.23.1rc0.</p></div>
    </content>
    <link href="https://cve.radiocsirt.org/vuln/cve-2026-54235"/>
  </entry>
  <entry>
    <id>https://cve.radiocsirt.org/vuln/ghsa-7h4p-rffg-7823</id>
    <title>GHSA-7h4p-rffg-7823 — vLLM: temperature=NaN and temperature=Infinity bypass validation and propagate to GPU kernels</title>
    <updated>2026-10-03T03:32:24.425785+00:00</updated>
    <content type="xhtml">
      <div xmlns="http://www.w3.org/1999/xhtml"><p><strong>Affected:</strong> PyPI: vllm</p>
<p>## Summary</p>
<p>All temperature validation gates use comparison operators (`&lt;`, `&gt;`), which silently evaluate to `False` for `NaN` and for positive `Infinity` in Python's IEEE 754 float semantics. Both values pass every guard and propagate to GPU sampling kernels, where they produce undefined behavior or CUDA errors that can crash the inference worker. Note: `-Infinity` is correctly caught.</p>
<p>## Root Cause</p>
<p>`sampling_params.py:384`:
```python
if 0 &lt; self.temperature &lt; _MAX_TEMP:  # NaN → False; +Inf → False
```</p>
<p>`sampling_params.py:462`:
```python
if self.temperature &lt; 0.0:            # NaN → False; +Inf → False
    raise VLLMValidationError(...)
```</p>
<p>No `math.isnan()` or `math.isinf()` check exists anywhere in `sampling_params.py`.</p>
<p>Python semantics (verified): `float('nan') &lt; 0.0` → `False`, `float('inf') &lt; 0.0` → `False`.</p>
<p>## Impact</p>
<p>Crash of inference worker on GPU kernel execution with NaN/Inf softmax input, degrading service for all concurrent users.</p>
<p>## Remediation</p>
<p>Add `math.isfinite(self.temperature)` check in `_verify_args()`. Reject non-finite float values with a 400 error.</p>
<p>## Fix</p>
<p>A fix for this vulnerability was merged here: https://github.com/vllm-project/vllm/pull/45116</p></div>
    </content>
    <link href="https://cve.radiocsirt.org/vuln/ghsa-7h4p-rffg-7823"/>
  </entry>
</feed>
