Topic: model

116 stories found

Today

newsletters48

[AINews] Qwen 3.8 Max(2.4T) and 27B, new open weights models for Coding and Cowork

Qwen 3.8, featuring the 2.4T version and a 27B model, has been released as new open-weight models for coding and collaborative work. This update marks an advancement in AI capabilities for programming tasks and enhanced teamwork functionalities.

latent.space
releases48

ggml/llama.cpp releases: b10251

The ggml/llama.cpp project released version b10251, which adds support for MTP in GLM-4.7-Flash (#24868). This update is significant as it enhances compatibility and functionality on Apple Silicon Macs, improving the performance of language models.

github.com
research40

Cost-Effective Automated Judging of Natural-Language Mathematical Proofs

A new method uses cheaper open-source language models to grade natural-language mathematical proofs, reducing costs associated with evaluating math-reasoning systems. This approach addresses the high expense of using advanced language models like LLMs for such tasks.

arxiv.org

Yesterday

ai_labs75

How we built a realtime system for responsive voice AI in six months

A team developed GPT-Live, a real-time voice AI system that allows for seamless, low-latency interactions, enabling more responsive and natural conversations within just six months. This achievement is significant as it enhances user experience in voice AI applications by reducing delays and improving the overall interaction quality.

openai.com
trending59

AirLLM 70B inference with single 4GB GPU

AirLLM 70B model can now be run using a single 4GB GPU, significantly reducing hardware requirements for large language models. This breakthrough could lower barriers to entry for deploying advanced AI models in resource-constrained environments.

github.com
research40

Can LLMs Really Understand Item Difficulty Levels? Implications for Automated Item Generation Using LLMs

The study examines whether large language models can accurately predict item difficulty, crucial for educational assessments, raising questions about their reliability in automated test generation.

arxiv.org

Saturday, August 1, 2026

industry24

AMD Releases Instella-MoE-16B-A3B: A Fully Open Mixture-of-Experts LLM With 2.8B Active Parameters Trained On Instinct GPUs

AMD has released Instella-MoE-16B-A3B, an open-source Mixture-of-Experts language model trained on Instinct GPUs, which could advance research by providing transparency in the training process.

marktechpost.com

Friday, July 31, 2026

research2 sources⚡ Corroborated40

BridgeAlign: Bridging Preference Alignment for Humanities and Social Sciences

BridgeAlign addresses the gap in data synthesis for large language models by focusing on humanities and social sciences, where nuanced quality judgments are crucial, rather than targeting domains with verifiable answers. This approach aims to improve the relevance and accuracy of LLMs in open-ended fields.

Covered by ArXiv cs.CL (Computation and Language / NLP)
releases48

ggml/llama.cpp releases: b10212

The ggml/llama.cpp project updated to load only necessary MTP tensors, reducing memory usage and improving efficiency for certain models. This update is significant as it optimizes performance without compromising functionality.

github.com
research40

Sympathetic Framing: Evaluating AI Alignment across Sociodemographic Groups

Researchers are examining whether large language models understand and convey emotional nuances through different sociodemographic frames, a critical aspect as these models increasingly influence public opinion. This study addresses concerns beyond bias, focusing on how LLMs align with sympathetic or empathetic framing across diverse groups.

arxiv.org

116 of 116 items shown. Sources: 76 days indexed.