Topic: model
116 stories found
Today

[AINews] Qwen 3.8 Max(2.4T) and 27B, new open weights models for Coding and Cowork
Qwen 3.8, featuring the 2.4T version and a 27B model, has been released as new open-weight models for coding and collaborative work. This update marks an advancement in AI capabilities for programming tasks and enhanced teamwork functionalities.
ggml/llama.cpp releases: b10251
The ggml/llama.cpp project released version b10251, which adds support for MTP in GLM-4.7-Flash (#24868). This update is significant as it enhances compatibility and functionality on Apple Silicon Macs, improving the performance of language models.
Cost-Effective Automated Judging of Natural-Language Mathematical Proofs
A new method uses cheaper open-source language models to grade natural-language mathematical proofs, reducing costs associated with evaluating math-reasoning systems. This approach addresses the high expense of using advanced language models like LLMs for such tasks.
Yesterday
How we built a realtime system for responsive voice AI in six months
A team developed GPT-Live, a real-time voice AI system that allows for seamless, low-latency interactions, enabling more responsive and natural conversations within just six months. This achievement is significant as it enhances user experience in voice AI applications by reducing delays and improving the overall interaction quality.
AirLLM 70B inference with single 4GB GPU
AirLLM 70B model can now be run using a single 4GB GPU, significantly reducing hardware requirements for large language models. This breakthrough could lower barriers to entry for deploying advanced AI models in resource-constrained environments.
Can LLMs Really Understand Item Difficulty Levels? Implications for Automated Item Generation Using LLMs
The study examines whether large language models can accurately predict item difficulty, crucial for educational assessments, raising questions about their reliability in automated test generation.
Saturday, August 1, 2026
AMD Releases Instella-MoE-16B-A3B: A Fully Open Mixture-of-Experts LLM With 2.8B Active Parameters Trained On Instinct GPUs
AMD has released Instella-MoE-16B-A3B, an open-source Mixture-of-Experts language model trained on Instinct GPUs, which could advance research by providing transparency in the training process.
Friday, July 31, 2026
BridgeAlign: Bridging Preference Alignment for Humanities and Social Sciences
BridgeAlign addresses the gap in data synthesis for large language models by focusing on humanities and social sciences, where nuanced quality judgments are crucial, rather than targeting domains with verifiable answers. This approach aims to improve the relevance and accuracy of LLMs in open-ended fields.
ggml/llama.cpp releases: b10212
The ggml/llama.cpp project updated to load only necessary MTP tensors, reducing memory usage and improving efficiency for certain models. This update is significant as it optimizes performance without compromising functionality.
Sympathetic Framing: Evaluating AI Alignment across Sociodemographic Groups
Researchers are examining whether large language models understand and convey emotional nuances through different sociodemographic frames, a critical aspect as these models increasingly influence public opinion. This study addresses concerns beyond bias, focusing on how LLMs align with sympathetic or empathetic framing across diverse groups.
116 of 116 items shown. Sources: 76 days indexed.