Topic: release
78 stories found
Today
ggml/llama.cpp releases: b10276
The ggml/llama.cpp project released version b10276, which includes a preference for using npm ci over install for enhanced security. This update is important as it addresses best practices in package management to reduce potential vulnerabilities.
Yesterday
ggml/llama.cpp releases: b10253
The ggml/llama.cpp project updated its cpp-httplib library to version 0.52.0 in commit b10253, which is significant for users of the macOS Apple Silicon version as it enhances their local AI model capabilities.
Ollama releases: v0.32.6
Ollama released version v0.32.6, which includes improvements to Qwen3.5's performance on Apple GPUs by automatically using the model's MTP head for speculative decoding and updates to `/v1/chat/completions` streaming to match OpenAI's wire format, enhancing compatibility.
Monday, August 3, 2026
NousResearch releases: Hermes Agent v0.20.0 (2026.8.3)
Hermes Agent v0.20.0 (v2026.8.3) was released on August 3, 2026, marking significant progress with over 3,650 commits and 1,400 merged pull requests, highlighting the project's active development and community contribution.
ggml/llama.cpp releases: b10236
The ggml/llama.cpp project has updated its codebase to include new Lightning Indexer implementations for both DSv4 and F16, enhancing support for specific input dimensions. These updates are crucial for improving performance in handling high-dimensional data with mixed precision, which is significant for advancing machine learning model efficiency.
Sunday, August 2, 2026
ggml/llama.cpp releases: b10235
The ggml/llama.cpp project has released version b10235, which includes the implementation of the SILU_BACK operation for f32 and fixes redundant asserts in the code. This update is significant as it enhances the functionality and stability of the machine learning library, particularly important for developers working on macOS Apple Silicon systems.
ggml/llama.cpp releases: b10232
The ggml/llama.cpp project released a new version implementing DeepSeek V4 hyper-connections, optimizing kernels for better performance on Metal devices. This update is significant as it enhances computational efficiency in deep learning models, particularly beneficial for hardware that supports Metal.
Saturday, August 1, 2026
ggml/llama.cpp releases: b10223
The ggml/llama.cpp project released version b10223 to fix CI errors. This update is important for ensuring the stability and reliability of the software, particularly on macOS Apple Silicon.
ggml/llama.cpp releases: b10219
The ggml/llama.cpp project updated its chat history feature to persist reasoning_content, addressing a previous limitation where only assistant content was stored, thus allowing better continuity of thought across interactions.
AMD Releases Instella-MoE-16B-A3B: A Fully Open Mixture-of-Experts LLM With 2.8B Active Parameters Trained On Instinct GPUs
AMD has released Instella-MoE-16B-A3B, an open-source Mixture-of-Experts language model trained on Instinct GPUs, which could advance research by providing transparency in the training process.
78 of 78 items shown. Sources: 77 days indexed.