CVE-2026-44223 - Vulnerability Details

- vLLM: extract_hidden_states speculative decoding crashes server on any request with penalty parameters

Description

vLLM is an inference and serving engine for large language models (LLMs). From to before 0.20.0, the extract_hidden_states speculative decoding proposer in vLLM returns a tensor with an incorrect shape after the first decode step, causing a RuntimeError that crashes the EngineCore process. The crash is triggered when any request in the batch uses sampling penalty parameters (repetition_penalty, frequency_penalty, or presence_penalty). A single request with a penalty parameter (e.g., "repetition_penalty": 1.1) is sufficient to crash the server. This vulnerability is fixed in 0.20.0.

Published: 2026-05-12

Score: 6.5 Medium

EPSS: < 1% Very Low

KEV: No

Impact:

Action:

Analysis

Analysis and contextual insights are available on OpenCVE Cloud.

Default status is the baseline for the product, each version can override it (e.g. patched versions marked unaffected).

Vendor Product Default status Versions

vllm-project

vllm

affected

Version	Status	Constraints
`>= 0.18.0, < 0.20.0`	affected	—

Configuration 1 [-]

cpe:2.3:a:vllm:vllm:*:*:*:*:*:*:*:*

No data.

Vendor Product Confidence Versions

Vllm-project

Vllm

100%

Version	Status	Scheme	Platform
`[0.18.0,0.20.0)`	affected	semver	—

Found an issue or want to improve our Enrichment? You can suggest it directly by opening an issue on our dedicated GitHub repository .

Remediation

No vendor fix or workaround currently provided.

Additional remediation guidance may be available on OpenCVE Cloud.

Tracking

Sign in to view the affected projects.

Advisories

Source	ID	Title
Github GHSA	GHSA-83vm-p52w-f9pw	vLLM: extract_hidden_states speculative decoding crashes server on any request with penalty parameters

No CVSS v4.0

Attack Vector Network

Attack Complexity Low

Privileges Required Low

Scope Unchanged

Confidentiality Impact None

Integrity Impact None

Availability Impact High

User Interaction None

No CVSS v3.0

No CVSS v2

This CVE is not in the KEV list.

The EPSS score is 0.00041.

Exploitation poc

Automatable no

Technical Impact partial

References

Link	Providers
https://github.com/vllm-project/vllm/pull/38610
https://github.com/vllm-project/vllm/security/advisories/GHSA-83vm-p52w-f9pw

History

Fri, 15 May 2026 15:15:00 +0000

Type	Values Removed	Values Added
Metrics		ssvc `{'options': {'Automatable': 'no', 'Exploitation': 'poc', 'Technical Impact': 'partial'}, 'version': '2.0.3'}`

Thu, 14 May 2026 15:45:00 +0000

Type	Values Removed	Values Added
First Time appeared		Vllm Vllm vllm
CPEs		cpe:2.3:a:vllm:vllm::::::::
Vendors & Products		Vllm Vllm vllm

Tue, 12 May 2026 23:30:00 +0000

Type	Values Removed	Values Added
First Time appeared		Vllm-project Vllm-project vllm
Vendors & Products		Vllm-project Vllm-project vllm

Tue, 12 May 2026 20:15:00 +0000

Type	Values Removed	Values Added
Description		vLLM is an inference and serving engine for large language models (LLMs). From to before 0.20.0, the extract_hidden_states speculative decoding proposer in vLLM returns a tensor with an incorrect shape after the first decode step, causing a RuntimeError that crashes the EngineCore process. The crash is triggered when any request in the batch uses sampling penalty parameters (repetition_penalty, frequency_penalty, or presence_penalty). A single request with a penalty parameter (e.g., "repetition_penalty": 1.1) is sufficient to crash the server. This vulnerability is fixed in 0.20.0.
Title		vLLM: extract_hidden_states speculative decoding crashes server on any request with penalty parameters
Weaknesses		CWE-131 CWE-704
References		https://github.com/vllm-project/vllm/pull/38610 https://github.com/vllm-project/vllm/security/advisories/GHSA-83vm-p52w-f9pw
Metrics		cvssV3_1 `{'score': 6.5, 'vector': 'CVSS:3.1/AV:N/AC:L/PR:L/UI:N/S:U/C:N/I:N/A:H'}`

Subscriptions

Vllm Vllm

Vllm-project Vllm

MITRE

Status: PUBLISHED

Assigner: GitHub_M

Published: 2026-05-12T19:58:40.862Z

Updated: 2026-05-15T14:46:25.695Z

Reserved: 2026-05-05T15:42:40.518Z

Link: CVE-2026-44223

Vulnrichment

Updated: 2026-05-15T14:43:40.735Z

NVD

Status : Modified

Published: 2026-05-12T20:16:43.293

Modified: 2026-05-15T15:16:52.560

Link: CVE-2026-44223

Redhat

No data.

OpenCVE Enrichment

Updated: 2026-05-12T23:15:26Z

Weaknesses

Tracking

Attack Vector Network

Attack Complexity Low

Privileges Required Low

Scope Unchanged

Confidentiality Impact None

Integrity Impact None

Availability Impact High

User Interaction None

Exploitation poc

Automatable no

Technical Impact partial

Subscriptions

JSON object

JSON object

JSON object

JSON object

JSON object