Количество 4
Количество 4
CVE-2026-90554
A flaw was found in vLLM. This vulnerability allows a remote attacker to cause a denial of service by providing a specially crafted, highly compressed video as multimodal input when the NanoNemotronVL model is configured to use audio from video. The lack of proper size or duration limits during audio extraction from the video input can force the server to allocate excessive memory, leading to system instability and preventing legitimate users from accessing the service.
CVE-2026-90554
vLLM versions >=0.10.2 and <0.28.0 do not apply any audio decode-size or duration limit when extracting audio from video input for NanoNemotronVL models. In nano_nemotron_vl.py, _extract_audio_from_videos calls load_audio_pyav(BytesIO(video_bytes)) without the max_duration_s or max_decode_bytes parameters, so neither VLLM_MAX_AUDIO_DECODE_DURATION_S nor VLLM_MAX_AUDIO_DECODE_BYTES is enforced (unlike the direct audio upload path in AudioMediaIO). When a NanoNemotronVL model is served with use_audio_in_video=True, an attacker who supplies a small, highly compressed video as multimodal input can force the server to allocate gigabytes of memory during audio decoding, resulting in a denial of service. Fixed in vLLM 0.28.0.
CVE-2026-90554
vLLM versions >=0.10.2 and <0.28.0 do not apply any audio decode-size ...
GHSA-m39m-2pqx-3vm2
vLLM versions >=0.10.2 and <0.28.0 do not apply any audio decode-size or duration limit when extracting audio from video input for NanoNemotronVL models. In nano_nemotron_vl.py, _extract_audio_from_videos calls load_audio_pyav(BytesIO(video_bytes)) without the max_duration_s or max_decode_bytes parameters, so neither VLLM_MAX_AUDIO_DECODE_DURATION_S nor VLLM_MAX_AUDIO_DECODE_BYTES is enforced (unlike the direct audio upload path in AudioMediaIO). When a NanoNemotronVL model is served with use_audio_in_video=True, an attacker who supplies a small, highly compressed video as multimodal input can force the server to allocate gigabytes of memory during audio decoding, resulting in a denial of service. Fixed in vLLM 0.28.0.
Уязвимостей на страницу
Уязвимость | CVSS | EPSS | Опубликовано | |
|---|---|---|---|---|
CVE-2026-90554 A flaw was found in vLLM. This vulnerability allows a remote attacker to cause a denial of service by providing a specially crafted, highly compressed video as multimodal input when the NanoNemotronVL model is configured to use audio from video. The lack of proper size or duration limits during audio extraction from the video input can force the server to allocate excessive memory, leading to system instability and preventing legitimate users from accessing the service. | CVSS3: 6.2 | 0% Низкий | 3 дня назад | |
CVE-2026-90554 vLLM versions >=0.10.2 and <0.28.0 do not apply any audio decode-size or duration limit when extracting audio from video input for NanoNemotronVL models. In nano_nemotron_vl.py, _extract_audio_from_videos calls load_audio_pyav(BytesIO(video_bytes)) without the max_duration_s or max_decode_bytes parameters, so neither VLLM_MAX_AUDIO_DECODE_DURATION_S nor VLLM_MAX_AUDIO_DECODE_BYTES is enforced (unlike the direct audio upload path in AudioMediaIO). When a NanoNemotronVL model is served with use_audio_in_video=True, an attacker who supplies a small, highly compressed video as multimodal input can force the server to allocate gigabytes of memory during audio decoding, resulting in a denial of service. Fixed in vLLM 0.28.0. | CVSS3: 6.2 | 0% Низкий | 3 дня назад | |
CVE-2026-90554 vLLM versions >=0.10.2 and <0.28.0 do not apply any audio decode-size ... | CVSS3: 6.2 | 0% Низкий | 3 дня назад | |
GHSA-m39m-2pqx-3vm2 vLLM versions >=0.10.2 and <0.28.0 do not apply any audio decode-size or duration limit when extracting audio from video input for NanoNemotronVL models. In nano_nemotron_vl.py, _extract_audio_from_videos calls load_audio_pyav(BytesIO(video_bytes)) without the max_duration_s or max_decode_bytes parameters, so neither VLLM_MAX_AUDIO_DECODE_DURATION_S nor VLLM_MAX_AUDIO_DECODE_BYTES is enforced (unlike the direct audio upload path in AudioMediaIO). When a NanoNemotronVL model is served with use_audio_in_video=True, an attacker who supplies a small, highly compressed video as multimodal input can force the server to allocate gigabytes of memory during audio decoding, resulting in a denial of service. Fixed in vLLM 0.28.0. | CVSS3: 6.2 | 0% Низкий | 3 дня назад |
Уязвимостей на страницу