CVE-2026-105755

CVE-2026-105755 published: vLLM is an inference and serving engine for large language models. Prior to 0.30.0, flash late-interaction scoring at the /score and /rerank endpoints derives each worker's query_key value from the caller-controlled X-Request-Id header. A concurrent request...

View full NVD advisory → ← Back to CVE watch