Topic: metal

14 stories found

Yesterday

releases48

ggml/llama.cpp releases: b10819

The ggml/llama.cpp project released a fix for a memory leak in early return functionality (commit b10819). This update is crucial as it addresses a potential stability issue, enhancing the reliability of the software.

github.com

Friday, September 4, 2026

releases48

ggml/llama.cpp releases: b10816

The ggml/llama.cpp project released a new version including tuning updates for the M3 model and additional precision settings, addressing formatting issues. These changes are significant for developers working with the M3 model to optimize performance on metal GPUs.

github.com

Wednesday, September 2, 2026

Tuesday, September 1, 2026

Monday, August 31, 2026

Sunday, August 30, 2026

Saturday, August 29, 2026

Wednesday, August 26, 2026

releases54

Ollama releases: v0.33.1

Ollama released version 0.33.1, which includes updates to Qwen3.8 Flash Next support, cmake patches, and mlxrunner structured output for improved model loading times, highlighting ongoing development and community contributions.

github.com

Monday, August 24, 2026

releases48

ggml/llama.cpp releases: b10615

The ggml/llama.cpp project has updated its Metal code to include device-tuned vectorized functions for flash attention, enhancing performance on specific hardware. These updates are crucial for optimizing the computational efficiency of large language models on Apple's M1 chips.

github.com

🌿 That's all for now. Come back tomorrow.

14 of 14 items shown. Sources: 107 days indexed.