Topic: metal

5 stories found

Today

releases48

ggml/llama.cpp releases: b11093

The ggml/llama.cpp project released a new version addressing an issue with mask bounds in the flash attention block pre-pass, aiming to improve performance and stability. This update is significant for users relying on macOS Apple Silicon (arm64) as it resolves specific technical problems affecting their systems.

github.comโ†—

Sunday, September 20, 2026

releases48

ggml/llama.cpp releases: b11064

The ggml/llama.cpp project updated its dsv4_hc_pre kernels to support arbitrary hardware contexts (hc), addressing a limitation that previously forced it to use CPU fallbacks. This change is significant as it enhances compatibility and performance across different hardware configurations, particularly for Kimi-K3 which uses varying hc values for banked checkpoints.

github.comโ†—

Friday, September 11, 2026

releases48

ggml/llama.cpp releases: b10909

The ggml/llama.cpp project updated its Metal backend to consolidate all fusable operation patterns into a single fusion table, enhancing debug capabilities. This change aims to improve optimization and maintainability of the codebase.

github.comโ†—

๐ŸŒฟ That's all for now. Come back tomorrow.

5 of 5 items shown. Sources: 123 days indexed.