Topic: split
4 stories found
Thursday, September 17, 2026
ggml/llama.cpp releases: b11011
The ggml/llama.cpp project released a fix for the function signature of `ggml_backend_sycl_split_buffer_type` to address compatibility issues, which is important for ensuring smooth operation on specific hardware platforms. This update affects users particularly interested in Apple Silicon Macs and iOS devices.
Wednesday, September 16, 2026
ggml/llama.cpp releases: b11009
The ggml/llama.cpp project released a new version addressing issues with split states and granularity for fused QKV gemm operations, crucial for models like Qwen35. This update ensures correct handling of attention layers when using specific configurations, enhancing the model's performance and accuracy.
Friday, September 11, 2026
ggml/llama.cpp releases: b10908
The ggml/llama.cpp project released a fix for idle threads in specific multiplication kernels used for neural networks with fewer than 1024 elements. This update generalizes a previous row split to additional related kernels, improving thread utilization and potentially enhancing performance.
Thursday, September 10, 2026
ggml/llama.cpp releases: b10899
The ggml/llama.cpp project released updates that optimize matrix multiplication operations for Vulkan, particularly focusing on improving performance with smaller matrices. These changes are significant as they enhance computational efficiency in models like Qwen, which can lead to faster processing times and better resource utilization.
🌿 That's all for now. Come back tomorrow.
4 of 4 items shown. Sources: 123 days indexed.