Topic: release

90 stories found

Today

releases2 sources⚡ Corroborated48

ggml/llama.cpp releases: b11094

The ggml/llama.cpp project updated the cpp-httplib library to version 0.57.1, a change signed off by Adrien Gallouët that enhances the project's functionality. This update is significant as it improves compatibility and performance for users of the llamacpp software suite.

Covered by ggml/llama.cpp releases
releases48

ggml/llama.cpp releases: b11095

The ggml/llama.cpp project has released updates focusing on HMX-optimized GATED_DELTA_NET, marking progress towards faster and more efficient processing. These developments are crucial as they aim to enhance the performance of large language models by optimizing hardware support.

github.com
releases42

vLLM releases: v0.30.0

vLLM released version 0.30.0 with significant contributions from many developers, introducing new models like DeepSeek-V4.1-Flash and DeepGEMM Mega-mHC, marking a substantial update in the model's capabilities.

github.com

Yesterday

releases60

NousResearch releases: Hermes Agent v0.21.4 (v2026.9.21)

Hermes Agent version v0.21.4 was released on September 21, 2026, consolidating approximately 1,800 merged pull requests to provide a stable release for various deployment environments, including Docker images and Hermes Cloud services.

github.com
releases48

ggml/llama.cpp releases: b11090

The ggml/llama.cpp project fixed a compilation error related to CUDA sm_70 tiles by generalizing the tile shape in version b11090, addressing an issue where a recent update mismatched tile definitions. This update is crucial for ensuring compatibility across different GPU architectures.

github.com

Sunday, September 20, 2026

releases2 sources⚡ Corroborated48

ggml/llama.cpp releases: b11065

The ggml/llama.cpp project released an update that tunes the FA parameter for use with the Gemma 4 on Ampere or newer GPUs, enhancing performance. This update is significant as it optimizes machine learning model processing for specific hardware, potentially improving speed and efficiency in applications like natural language processing.

Covered by ggml/llama.cpp releases
releases48

ggml/llama.cpp releases: b11064

The ggml/llama.cpp project updated its dsv4_hc_pre kernels to support arbitrary hardware contexts (hc), addressing a limitation that previously forced it to use CPU fallbacks. This change is significant as it enhances compatibility and performance across different hardware configurations, particularly for Kimi-K3 which uses varying hc values for banked checkpoints.

github.com

Saturday, September 19, 2026

releases2 sources⚡ Corroborated54

Ollama releases: v0.34.3

Ollama released version v0.34.3, which includes updates to the `GET /api/show` endpoint to advertise each model's thinking controls and default settings, enhancing API transparency. This update is significant for users needing detailed control over model behavior.

Covered by Ollama releases, ggml/llama.cpp releases
releases2 sources⚡ Corroborated48

ggml/llama.cpp releases: b11056

The ggml/llama.cpp project released version b11056, which includes a change to enable I32 GET_ROWS (#29116). This update is significant for improving the handling of 32-bit integer operations in the library, potentially enhancing performance and functionality.

Covered by ggml/llama.cpp releases
releases48

ggml/llama.cpp releases: b11054

The ggml/llama.cpp project updated to enable support for the TOP_K operation on the Hexagon processor, improving row partitioning and optimizing large-row selection. This update is crucial as it enhances the efficiency of operations involving top-k elements in machine learning models running on Hexagon hardware.

github.com

90 of 90 items shown. Sources: 123 days indexed.