Topic: support

25 stories found

Today

releases48

ggml/llama.cpp releases: b11095

The ggml/llama.cpp project has released updates focusing on HMX-optimized GATED_DELTA_NET, marking progress towards faster and more efficient processing. These developments are crucial as they aim to enhance the performance of large language models by optimizing hardware support.

github.com

Yesterday

research40

SAGE: Schema-Guided LLMs for Grant Review

SAGE, a schema-guided language model, assists grant reviewers by systematically evaluating applications based on detailed criteria, potentially improving the consistency and transparency of the review process.

arxiv.org
research35

LoRA Enhanced Contrastive Learning with SAS Vision Transformers

A new method called LoRA Enhanced Contrastive Learning with SAS Vision Transformers has been developed to improve automatic target recognition in synthetic aperture sonar, addressing limitations such as sparse target data and background noise. This advancement is crucial for enhancing naval capabilities through more efficient and accurate underwater target identification.

arxiv.org

Sunday, September 20, 2026

open_source62

langgenius/dify — Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cl

Agentic offers a collaborative workspace for building AI workflows and RAG pipelines with extensive model and tool support, enabling seamless deployment options from cloud to self-hosted environments, facilitating rapid transition from prototypes to production.

github.com
releases48

ggml/llama.cpp releases: b11064

The ggml/llama.cpp project updated its dsv4_hc_pre kernels to support arbitrary hardware contexts (hc), addressing a limitation that previously forced it to use CPU fallbacks. This change is significant as it enhances compatibility and performance across different hardware configurations, particularly for Kimi-K3 which uses varying hc values for banked checkpoints.

github.com

Saturday, September 19, 2026

releases48

ggml/llama.cpp releases: b11055

The ggml/llama.cpp project released version b11055, which includes support for GEGLU_QUICK in hexagon models. This update enhances the model's performance by optimizing certain computations, making it more efficient and potentially improving user experience on compatible devices.

github.com

Friday, September 18, 2026

releases3 sources🛡️ Verified48

ggml/llama.cpp releases: b11046

The ggml/llama.cpp project has released an update that adds support for the `flash_attn_f32_f16_bin` kernel in OpenCL, enhancing computational efficiency for certain operations. This update is significant as it improves performance in processing tasks related to large language models.

Covered by ggml/llama.cpp releases

Thursday, September 17, 2026

releases48

ggml/llama.cpp releases: b11025

The ggml/llama.cpp project released a new version that includes a fix for extending Nemotron MTP support, removing unnecessary declarations. This update is important as it enhances the model's compatibility and efficiency.

github.com
research40

Think Before You Comfort: Reflective Cognitive Alignment for Protocol-Grounded Elderly Stimulation Agents

A new protocol aims to enhance the scalability of Cognitive Stimulation Therapy (CST) for elderly individuals with cognitive impairments by reducing dependency on trained facilitators and addressing data scarcity issues. This advancement is crucial as it could broaden access to non-pharmacological support while respecting privacy concerns.

arxiv.org

Wednesday, September 16, 2026

releases48

ggml/llama.cpp releases: b11006

The ggml/llama.cpp project released a new version that includes support for K-Quants Q4_K and Q6_K, enhancing quantization techniques crucial for optimizing machine learning models' performance on resource-constrained devices. This update is significant as it improves model efficiency without compromising too much on accuracy, making it more viable for deployment in edge computing scenarios.

github.com

25 of 25 items shown. Sources: 123 days indexed.