Topic: llms
36 stories found
Today
Token Signatures of Code: Comparing Coding Behaviors Across Large Language Models
A new study evaluates large language models (LLMs) based on their coding behaviors rather than just performance metrics like pass@k, highlighting that as models improve, traditional evaluation methods become less effective in distinguishing between them.
Yesterday

Pruning LLMs Like a Physicist: Block Removal as an Ising Optimization Problem
Researchers propose pruning large language models (LLMs) using techniques inspired by physics, specifically the Ising model for optimization problems, to improve efficiency without significantly impacting performance. This approach could lead to more resource-efficient LLMs, advancing practical applications and reducing computational costs.
From Discharge Notes to Patient Understanding: Persona-Grounded, Open-Ended Simulation of LLMs as Discharge Educators
A new study proposes using large language models (LLMs) in a persona-grounded, open-ended simulation as discharge educators to better adapt to patients' literacy, recall, and personality needs, addressing limitations of current LLM evaluations that focus on static or artifact-generation tasks. This approach aims to improve patient understanding and adherence to discharge plans.
Friday, September 18, 2026
Cache-to-Cache: Direct Semantic Communication Between LLMs (2025)
Researchers have developed a method called "Cache-to-Cache" that allows large language models to communicate directly, enhancing collaboration and potentially improving model performance. This breakthrough could significantly advance the field of artificial intelligence by enabling more efficient and effective information sharing among different language models.
How OpenAI Used Its Own LLMs to Design Its Jalapeño Chip
OpenAI utilized its large language models (LLMs) to design the Jalapeño chip, a significant step as it demonstrates AI's potential in hardware development and could pave the way for more efficient and automated chip design processes.
Neo-Classic: A Benchmark for Evaluating Linguistic-Aesthetic Reasoning in Classical Chinese Poetry
A new benchmark called Neo-Classic has been developed to evaluate the ability of models to demonstrate true linguistic-aesthetic reasoning in classical Chinese poetry, rather than relying on memorized patterns. This is important because it helps distinguish between superficial accuracy and deeper understanding in AI models.
Thursday, September 17, 2026

I Don't Like LLMs
The article expresses dissatisfaction with large language models, highlighting concerns about their limitations and potential biases. This matter is significant as LLMs are increasingly integrated into various applications, making public opinion crucial for their development and use.
Faking Good and Faking Bad in LLMs: Response Distortion Across Dark Triad Personality Traits
The study explores how large language models (LLMs) are influenced by social desirability and impression management, similar to humans during personality assessments, highlighting the need for better understanding of response distortions in AI.
Wednesday, September 16, 2026

PS5 Linux lead quits: "a bunch of noobs using LLMs" that "they don't understand"
A developer who led the effort to create a Linux version for the PlayStation 5 has quit, criticizing those attempting to use such tools without understanding them. This departure highlights ongoing challenges in cross-platform development and user knowledge gaps.
[AINews] Jev: a “System One Model” that only decides/classifies/routes/scores — >100x faster, >200x cheaper than small frontier LLMs
Jev introduces a "System One Model" for decision-making processes that is reportedly 100 times faster and over 200 times cheaper than smaller language models. This development could significantly impact cost-efficiency in various industries by enhancing speed and reducing expenses associated with artificial intelligence solutions.
36 of 36 items shown. Sources: 123 days indexed.