Topic: hexagon

9 stories found

Today

releases48

ggml/llama.cpp releases: b11095

The ggml/llama.cpp project has released updates focusing on HMX-optimized GATED_DELTA_NET, marking progress towards faster and more efficient processing. These developments are crucial as they aim to enhance the performance of large language models by optimizing hardware support.

github.com↗

Saturday, September 19, 2026

releases2 sources⚔ Corroborated48

ggml/llama.cpp releases: b11056

The ggml/llama.cpp project released version b11056, which includes a change to enable I32 GET_ROWS (#29116). This update is significant for improving the handling of 32-bit integer operations in the library, potentially enhancing performance and functionality.

Covered by ggml/llama.cpp releases
releases48

ggml/llama.cpp releases: b11054

The ggml/llama.cpp project updated to enable support for the TOP_K operation on the Hexagon processor, improving row partitioning and optimizing large-row selection. This update is crucial as it enhances the efficiency of operations involving top-k elements in machine learning models running on Hexagon hardware.

github.com↗

Friday, September 18, 2026

releases2 sources⚔ Corroborated48

ggml/llama.cpp releases: b11045

The ggml/llama.cpp project released version b11045, which includes support for the ROLL operation. This update is significant as it enhances the project's capabilities, particularly in handling specific types of data transformations.

Covered by ggml/llama.cpp releases
releases48

ggml/llama.cpp releases: b11044

The ggml/llama.cpp project released an update that includes improvements to the hexagon backend, specifically enhancing IM2COL operations for both 1D and padded inputs. These changes are significant as they optimize memory access patterns, potentially improving performance in certain machine learning tasks.

github.com↗

Wednesday, September 16, 2026

releases48

ggml/llama.cpp releases: b11006

The ggml/llama.cpp project released a new version that includes support for K-Quants Q4_K and Q6_K, enhancing quantization techniques crucial for optimizing machine learning models' performance on resource-constrained devices. This update is significant as it improves model efficiency without compromising too much on accuracy, making it more viable for deployment in edge computing scenarios.

github.com↗

Tuesday, September 15, 2026

releases48

ggml/llama.cpp releases: b10991

The ggml/llama.cpp project released version b10991, which includes the addition of a missing contiguous fast-path and hvx_copy_uu for each run. This update is significant as it enhances performance on specific hardware configurations.

github.com↗

🌿 That's all for now. Come back tomorrow.

9 of 9 items shown. Sources: 123 days indexed.