Topic: hexagon
9 stories found
Today
ggml/llama.cpp releases: b11095
The ggml/llama.cpp project has released updates focusing on HMX-optimized GATED_DELTA_NET, marking progress towards faster and more efficient processing. These developments are crucial as they aim to enhance the performance of large language models by optimizing hardware support.
Saturday, September 19, 2026
ggml/llama.cpp releases: b11056
The ggml/llama.cpp project released version b11056, which includes a change to enable I32 GET_ROWS (#29116). This update is significant for improving the handling of 32-bit integer operations in the library, potentially enhancing performance and functionality.
ggml/llama.cpp releases: b11054
The ggml/llama.cpp project updated to enable support for the TOP_K operation on the Hexagon processor, improving row partitioning and optimizing large-row selection. This update is crucial as it enhances the efficiency of operations involving top-k elements in machine learning models running on Hexagon hardware.
Friday, September 18, 2026
ggml/llama.cpp releases: b11045
The ggml/llama.cpp project released version b11045, which includes support for the ROLL operation. This update is significant as it enhances the project's capabilities, particularly in handling specific types of data transformations.
ggml/llama.cpp releases: b11044
The ggml/llama.cpp project released an update that includes improvements to the hexagon backend, specifically enhancing IM2COL operations for both 1D and padded inputs. These changes are significant as they optimize memory access patterns, potentially improving performance in certain machine learning tasks.
Wednesday, September 16, 2026
ggml/llama.cpp releases: b11006
The ggml/llama.cpp project released a new version that includes support for K-Quants Q4_K and Q6_K, enhancing quantization techniques crucial for optimizing machine learning models' performance on resource-constrained devices. This update is significant as it improves model efficiency without compromising too much on accuracy, making it more viable for deployment in edge computing scenarios.
Tuesday, September 15, 2026
ggml/llama.cpp releases: b10991
The ggml/llama.cpp project released version b10991, which includes the addition of a missing contiguous fast-path and hvx_copy_uu for each run. This update is significant as it enhances performance on specific hardware configurations.
šæ That's all for now. Come back tomorrow.
9 of 9 items shown. Sources: 123 days indexed.