Topic: github
41 stories found
Today
ggml/llama.cpp releases: b11094
The ggml/llama.cpp project updated the cpp-httplib library to version 0.57.1, a change signed off by Adrien Gallouët that enhances the project's functionality. This update is significant as it improves compatibility and performance for users of the llamacpp software suite.
Yesterday
ggml/llama.cpp releases: b11090
The ggml/llama.cpp project fixed a compilation error related to CUDA sm_70 tiles by generalizing the tile shape in version b11090, addressing an issue where a recent update mismatched tile definitions. This update is crucial for ensuring compatibility across different GPU architectures.
Sunday, September 20, 2026
ggml/llama.cpp releases: b11065
The ggml/llama.cpp project released an update that tunes the FA parameter for use with the Gemma 4 on Ampere or newer GPUs, enhancing performance. This update is significant as it optimizes machine learning model processing for specific hardware, potentially improving speed and efficiency in applications like natural language processing.
Saturday, September 19, 2026
ggml/llama.cpp releases: b11056
The ggml/llama.cpp project released version b11056, which includes a change to enable I32 GET_ROWS (#29116). This update is significant for improving the handling of 32-bit integer operations in the library, potentially enhancing performance and functionality.
Friday, September 18, 2026
ggml/llama.cpp releases: b11046
The ggml/llama.cpp project has released an update that adds support for the `flash_attn_f32_f16_bin` kernel in OpenCL, enhancing computational efficiency for certain operations. This update is significant as it improves performance in processing tasks related to large language models.
Thursday, September 17, 2026
ggml/llama.cpp releases: b11028
The ggml/llama.cpp project released version b11028, which includes a fix for evicting old files. This update is important as it enhances the project's stability and efficiency, particularly for users on Apple Silicon platforms.
ggml/llama.cpp releases: b11011
The ggml/llama.cpp project released a fix for the function signature of `ggml_backend_sycl_split_buffer_type` to address compatibility issues, which is important for ensuring smooth operation on specific hardware platforms. This update affects users particularly interested in Apple Silicon Macs and iOS devices.
Wednesday, September 16, 2026
ggml/llama.cpp releases: b11007
The ggml/llama.cpp project released a new version that enables CUDA graph usage for MTP, improving performance. This update is significant as it enhances the efficiency of the software, particularly relevant for users requiring high computational power.
Tuesday, September 15, 2026
Ollama releases: v0.34.2
Ollama released version 0.34.2, which includes updates to llama.cpp; this update is significant for users relying on the software's text generation capabilities.
ggml/llama.cpp releases: b10991
The ggml/llama.cpp project released version b10991, which includes the addition of a missing contiguous fast-path and hvx_copy_uu for each run. This update is significant as it enhances performance on specific hardware configurations.
41 of 41 items shown. Sources: 123 days indexed.