Topic: x

222 stories found

Today

releases48

ggml/llama.cpp releases: b11095

The ggml/llama.cpp project has released updates focusing on HMX-optimized GATED_DELTA_NET, marking progress towards faster and more efficient processing. These developments are crucial as they aim to enhance the performance of large language models by optimizing hardware support.

github.com
releases42

vLLM releases: v0.30.0

vLLM released version 0.30.0 with significant contributions from many developers, introducing new models like DeepSeek-V4.1-Flash and DeepGEMM Mega-mHC, marking a substantial update in the model's capabilities.

github.com
research40

Recognition, Simulation, and Refusal: A Contamination-Aware Study of Classic Psychological Effects in LLM Agents

The study PsyAgentBench re-runs classic psychology experiments on LLMs to assess their susceptibility to human biases without attributing those biases directly to the models, highlighting the need for contamination-aware analysis. This matters as it provides a framework to understand and mitigate potential psychological effect mimicry in AI systems.

arxiv.org

Yesterday

releases2 sources⚡ Corroborated48

ggml/llama.cpp releases: b11090

The ggml/llama.cpp project fixed a compilation error related to CUDA sm_70 tiles by generalizing the tile shape in version b11090, addressing an issue where a recent update mismatched tile definitions. This update is crucial for ensuring compatibility across different GPU architectures.

Covered by ggml/llama.cpp releases, ArXiv cs.CL (Computation and Language / NLP)
releases2 sources⚡ Corroborated48

ggml/llama.cpp releases: b11081

The ggml/llama.cpp project released a new version with updates to make tensor data standard deviation configurable and added more examples to the documentation, enhancing flexibility and usability for developers. These changes are significant as they improve the toolkit's adaptability and user-friendliness in handling large language model architectures.

Covered by ggml/llama.cpp releases, ArXiv cs.CL (Computation and Language / NLP)
ai_labs75

Building standards for the next phase of AI

OpenAI proposes establishing global standards for AI to enhance safety through coordinated evaluation and governance. This initiative aims to address AI risks by fostering international cooperation in the tech sector.

openai.com
research40

Do small language models know what they don't know?

Researchers investigated ways to enhance the accuracy of small language models (with less than 3 billion parameters) using entropy-based confidence signals, finding potential improvements for models running on consumer hardware. This matters because it could make advanced language capabilities more accessible on standard devices.

arxiv.org
research35

LoRA Enhanced Contrastive Learning with SAS Vision Transformers

A new method called LoRA Enhanced Contrastive Learning with SAS Vision Transformers has been developed to improve automatic target recognition in synthetic aperture sonar, addressing limitations such as sparse target data and background noise. This advancement is crucial for enhancing naval capabilities through more efficient and accurate underwater target identification.

arxiv.org

Sunday, September 20, 2026

releases48

ggml/llama.cpp releases: b11064

The ggml/llama.cpp project updated its dsv4_hc_pre kernels to support arbitrary hardware contexts (hc), addressing a limitation that previously forced it to use CPU fallbacks. This change is significant as it enhances compatibility and performance across different hardware configurations, particularly for Kimi-K3 which uses varying hc values for banked checkpoints.

github.com

Saturday, September 19, 2026

releases2 sources⚡ Corroborated48

ggml/llama.cpp releases: b11056

The ggml/llama.cpp project released version b11056, which includes a change to enable I32 GET_ROWS (#29116). This update is significant for improving the handling of 32-bit integer operations in the library, potentially enhancing performance and functionality.

Covered by ggml/llama.cpp releases

222 of 222 items shown. Sources: 123 days indexed.