Topic: authored

6 stories found

Friday, August 28, 2026

releases48

ggml/llama.cpp releases: b10665

The ggml/llama.cpp project added DSpark support for Nemotron3.5 in version b10665, enhancing model compatibility and functionality. This update is significant as it broadens the software's applicability for users with specific hardware configurations.

github.com

Thursday, August 27, 2026

releases48

ggml/llama.cpp releases: b10662

The ggml/llama.cpp project updated to include a new `--kv-unified-per-slot` argument for managing context pools in the KV cache, aiming to optimize memory usage and performance. This update is significant as it enhances the flexibility and efficiency of the model's context handling, which could lead to better resource management during large-scale language processing tasks.

github.com

Tuesday, August 25, 2026

releases48

ggml/llama.cpp releases: b10628

The ggml/llama.cpp project added support for Apple Remote Direct Memory Access (RDMA) as an RPC transport, enhancing network communication efficiency. This update is significant as it improves the project's compatibility with Apple hardware, potentially boosting performance in distributed computing scenarios.

github.com

Monday, August 24, 2026

releases48

ggml/llama.cpp releases: b10612

The ggml/llama.cpp project released version b10612, which disables a specific test for WebGPU compatibility. This update is important as it enhances the project's support for diverse hardware architectures, including Apple Silicon on macOS.

github.com

🌿 That's all for now. Come back tomorrow.

6 of 6 items shown. Sources: 107 days indexed.