← Vulnerability feed

Vulnerability record · CVE-2026-33298 · published 24 March 2026

CVE-2026-33298: Ggml llama.cpp heap-based buffer overflow vulnerability

Ggml · Llama.Cpp

llama.cpp is an inference of several LLM models in C/C++. Prior to b7824, an integer overflow vulnerability in the `ggml_nbytes` function allows an attacker to bypass memory validation by crafting a GGUF file with specific tensor dimensions. This causes `ggml_nbytes` to return a significantly smaller size than required (e.g., 4MB instead of Exabytes), leading to a heap-based buffer overflow when the application subsequently processes the tensor. This vulnerability allows potential Remote Code Execution (RCE) via memory corruption. b7824 contains a fix.

7.8 CVSS 3.1 High EPSS 0.37% · top 71.7% CWE-122 · Heap-based buffer overflowCWE-190 · Integer overflow
7.8CVSS 3.1 base score
0.37%EPSS exploitation probability, 30 days
NoNot in CISA KEV
1Affected product versions listed by NVD
2References, 1 tagged exploit
17 Jun 2026Last modified by NVD

Description

llama.cpp is an inference of several LLM models in C/C++. Prior to b7824, an integer overflow vulnerability in the `ggml_nbytes` function allows an attacker to bypass memory validation by crafting a GGUF file with specific tensor dimensions. This causes `ggml_nbytes` to return a significantly smaller size than required (e.g., 4MB instead of Exabytes), leading to a heap-based buffer overflow when the application subsequently processes the tensor. This vulnerability allows potential Remote Code Execution (RCE) via memory corruption. b7824 contains a fix.

CVSS:3.1/AV:L/AC:L/PR:N/UI:R/S:U/C:H/I:H/A:H

Affected products

1 vulnerable configurations from NVD's CPE data, grouped by vendor and product.

References

Track CVE-2026-33298 inside VULONE

Watch it alongside the ransomware crews, C2 infrastructure and forum chatter that reference it, query it through the API and pull it into your SIEM over TAXII.

Start free Open in platform

Related vulnerabilities

Same products first, then exploited flaws of the same weakness class.

9.8CVE-2026-34159Ggml llama.cpp memory buffer overflow vulnerabilityllama.cpp is an inference of several LLM models in C/C++. Prior to version b8492, the RPC backend's deserialize_tensor() skips all bounds validation …EPSS 1.2%9.8CVE-2026-21869Ggml llama.cpp out-of-bounds write vulnerabilityllama.cpp is an inference of several LLM models in C/C++. In commits 55d4206c8 and prior, the n_discard parameter is parsed directly from JSON input …EPSS 0.52%9.8CVE-2024-42478Ggml llama.cpp out-of-bounds read vulnerabilityllama.cpp provides LLM inference in C/C++. The unsafe `data` pointer member in the `rpc_tensor` structure can cause arbitrary address reading. This v…EPSS 0.60%9.8CVE-2024-42479Ggml llama.cpp out-of-bounds write vulnerabilityllama.cpp provides LLM inference in C/C++. The unsafe `data` pointer member in the `rpc_tensor` structure can cause arbitrary address writing. This v…EPSS 2.6%9.8CVE-2024-23605Ggml llama.cpp integer overflow vulnerabilityA heap-based buffer overflow vulnerability exists in the GGUF library header.n_kv functionality of llama.cpp Commit 18c2e17. A specially crafted .ggu…EPSS 1.3%9.8CVE-2024-23496Ggml llama.cpp integer overflow vulnerabilityA heap-based buffer overflow vulnerability exists in the GGUF library gguf_fread_str functionality of llama.cpp Commit 18c2e17. A specially crafted .…EPSS 1.3%9.8CVE-2024-21802Ggml llama.cpp heap-based buffer overflow vulnerabilityA heap-based buffer overflow vulnerability exists in the GGUF library info->ne functionality of llama.cpp Commit 18c2e17. A specially crafted .ggu…EPSS 1.4%9.8CVE-2024-21825Ggml llama.cpp integer overflow vulnerabilityA heap-based buffer overflow vulnerability exists in the GGUF library GGUF_TYPE_ARRAY/GGUF_TYPE_STRING parsing functionality of llama.cpp Commit 18c2…EPSS 1.3%

Source: NIST National Vulnerability Database (record CVE-2026-33298), CISA KEV, FIRST EPSS (scores of 2026-09-26). This page is refreshed as NVD updates the record.