Topic: kernels
3 stories found
Sunday, September 20, 2026
ggml/llama.cpp releases: b11064
The ggml/llama.cpp project updated its dsv4_hc_pre kernels to support arbitrary hardware contexts (hc), addressing a limitation that previously forced it to use CPU fallbacks. This change is significant as it enhances compatibility and performance across different hardware configurations, particularly for Kimi-K3 which uses varying hc values for banked checkpoints.
Wednesday, September 16, 2026
ggml/llama.cpp releases: b11006
The ggml/llama.cpp project released a new version that includes support for K-Quants Q4_K and Q6_K, enhancing quantization techniques crucial for optimizing machine learning models' performance on resource-constrained devices. This update is significant as it improves model efficiency without compromising too much on accuracy, making it more viable for deployment in edge computing scenarios.
Friday, September 11, 2026
ggml/llama.cpp releases: b10908
The ggml/llama.cpp project released a fix for idle threads in specific multiplication kernels used for neural networks with fewer than 1024 elements. This update generalizes a previous row split to additional related kernels, improving thread utilization and potentially enhancing performance.
🌿 That's all for now. Come back tomorrow.
3 of 3 items shown. Sources: 123 days indexed.