Skip to content

Commit b576c54

Browse files
committed
feat: update llama.cpp to 78d2f5246
1 parent 346853c commit b576c54

3 files changed

Lines changed: 19 additions & 1 deletion

File tree

‎CHANGELOG.md‎

Lines changed: 2 additions & 0 deletions
Original file line numberDiff line numberDiff line change
@@ -7,6 +7,8 @@ and this project adheres to [Semantic Versioning](https://semver.org/spec/v2.0.0
77

88
## [Unreleased]
99

10+
- feat: update llama.cpp to ggml-org/llama.cpp@78d2f5246
11+
1012
## [0.3.32]
1113

1214
- feat(example): support chained NextN heads for server MTP drafting

‎llama_cpp/llama_cpp.py‎

Lines changed: 16 additions & 0 deletions
Original file line numberDiff line numberDiff line change
@@ -1272,6 +1272,14 @@ def llama_flash_attn_type_name(flash_attn_type: int, /) -> Optional[bytes]:
12721272
...
12731273

12741274

1275+
# // Get the model file type (quantization) as a string, e.g. "Q8_0" or "Q4_K - Medium"
1276+
# LLAMA_API const char * llama_ftype_name(enum llama_ftype ftype);
1277+
@ctypes_function("llama_ftype_name", [ctypes.c_int], ctypes.c_char_p)
1278+
def llama_ftype_name(ftype: int, /) -> Optional[bytes]:
1279+
"""Get the model file type as a string."""
1280+
...
1281+
1282+
12751283
# // Initialize the llama + ggml backend
12761284
# // If numa is true, use NUMA optimizations
12771285
# // Call once at the start of the program
@@ -1910,6 +1918,14 @@ def llama_model_desc(
19101918
...
19111919

19121920

1921+
# // Get the model file type (quantization), e.g. LLAMA_FTYPE_MOSTLY_Q8_0
1922+
# LLAMA_API enum llama_ftype llama_model_ftype(const struct llama_model * model);
1923+
@ctypes_function("llama_model_ftype", [llama_model_p_ctypes], ctypes.c_int)
1924+
def llama_model_ftype(model: llama_model_p, /) -> int:
1925+
"""Get the model file type."""
1926+
...
1927+
1928+
19131929
# // Returns the total size of all the tensors in the model in bytes
19141930
# LLAMA_API uint64_t llama_model_size(const struct llama_model * model);
19151931
@ctypes_function("llama_model_size", [llama_model_p_ctypes], ctypes.c_uint64)

‎vendor/llama.cpp‎

Submodule llama.cpp updated 163 files

0 commit comments

Comments
 (0)