Topic: support
25 stories found
Today
ggml/llama.cpp releases: b11095
The ggml/llama.cpp project has released updates focusing on HMX-optimized GATED_DELTA_NET, marking progress towards faster and more efficient processing. These developments are crucial as they aim to enhance the performance of large language models by optimizing hardware support.
Yesterday
SAGE: Schema-Guided LLMs for Grant Review
SAGE, a schema-guided language model, assists grant reviewers by systematically evaluating applications based on detailed criteria, potentially improving the consistency and transparency of the review process.
LoRA Enhanced Contrastive Learning with SAS Vision Transformers
A new method called LoRA Enhanced Contrastive Learning with SAS Vision Transformers has been developed to improve automatic target recognition in synthetic aperture sonar, addressing limitations such as sparse target data and background noise. This advancement is crucial for enhancing naval capabilities through more efficient and accurate underwater target identification.
Sunday, September 20, 2026
langgenius/dify β Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cl
Agentic offers a collaborative workspace for building AI workflows and RAG pipelines with extensive model and tool support, enabling seamless deployment options from cloud to self-hosted environments, facilitating rapid transition from prototypes to production.
ggml/llama.cpp releases: b11064
The ggml/llama.cpp project updated its dsv4_hc_pre kernels to support arbitrary hardware contexts (hc), addressing a limitation that previously forced it to use CPU fallbacks. This change is significant as it enhances compatibility and performance across different hardware configurations, particularly for Kimi-K3 which uses varying hc values for banked checkpoints.
Saturday, September 19, 2026
ggml/llama.cpp releases: b11055
The ggml/llama.cpp project released version b11055, which includes support for GEGLU_QUICK in hexagon models. This update enhances the model's performance by optimizing certain computations, making it more efficient and potentially improving user experience on compatible devices.
Friday, September 18, 2026
ggml/llama.cpp releases: b11046
The ggml/llama.cpp project has released an update that adds support for the `flash_attn_f32_f16_bin` kernel in OpenCL, enhancing computational efficiency for certain operations. This update is significant as it improves performance in processing tasks related to large language models.
Thursday, September 17, 2026
ggml/llama.cpp releases: b11025
The ggml/llama.cpp project released a new version that includes a fix for extending Nemotron MTP support, removing unnecessary declarations. This update is important as it enhances the model's compatibility and efficiency.
Think Before You Comfort: Reflective Cognitive Alignment for Protocol-Grounded Elderly Stimulation Agents
A new protocol aims to enhance the scalability of Cognitive Stimulation Therapy (CST) for elderly individuals with cognitive impairments by reducing dependency on trained facilitators and addressing data scarcity issues. This advancement is crucial as it could broaden access to non-pharmacological support while respecting privacy concerns.
Wednesday, September 16, 2026
ggml/llama.cpp releases: b11006
The ggml/llama.cpp project released a new version that includes support for K-Quants Q4_K and Q6_K, enhancing quantization techniques crucial for optimizing machine learning models' performance on resource-constrained devices. This update is significant as it improves model efficiency without compromising too much on accuracy, making it more viable for deployment in edge computing scenarios.
25 of 25 items shown. Sources: 123 days indexed.