Topic: llms

36 stories found

Today

research40

Token Signatures of Code: Comparing Coding Behaviors Across Large Language Models

A new study evaluates large language models (LLMs) based on their coding behaviors rather than just performance metrics like pass@k, highlighting that as models improve, traditional evaluation methods become less effective in distinguishing between them.

arxiv.org

Yesterday

ai_labs67

Pruning LLMs Like a Physicist: Block Removal as an Ising Optimization Problem

Researchers propose pruning large language models (LLMs) using techniques inspired by physics, specifically the Ising model for optimization problems, to improve efficiency without significantly impacting performance. This approach could lead to more resource-efficient LLMs, advancing practical applications and reducing computational costs.

huggingface.co
research40

From Discharge Notes to Patient Understanding: Persona-Grounded, Open-Ended Simulation of LLMs as Discharge Educators

A new study proposes using large language models (LLMs) in a persona-grounded, open-ended simulation as discharge educators to better adapt to patients' literacy, recall, and personality needs, addressing limitations of current LLM evaluations that focus on static or artifact-generation tasks. This approach aims to improve patient understanding and adherence to discharge plans.

arxiv.org

Friday, September 18, 2026

trending2 sources⚡ Corroborated46

Cache-to-Cache: Direct Semantic Communication Between LLMs (2025)

Researchers have developed a method called "Cache-to-Cache" that allows large language models to communicate directly, enhancing collaboration and potentially improving model performance. This breakthrough could significantly advance the field of artificial intelligence by enabling more efficient and effective information sharing among different language models.

Covered by HN top LLM 24h, ArXiv cs.CL (Computation and Language / NLP)
trending50

How OpenAI Used Its Own LLMs to Design Its Jalapeño Chip

OpenAI utilized its large language models (LLMs) to design the Jalapeño chip, a significant step as it demonstrates AI's potential in hardware development and could pave the way for more efficient and automated chip design processes.

spectrum.ieee.org
research40

Neo-Classic: A Benchmark for Evaluating Linguistic-Aesthetic Reasoning in Classical Chinese Poetry

A new benchmark called Neo-Classic has been developed to evaluate the ability of models to demonstrate true linguistic-aesthetic reasoning in classical Chinese poetry, rather than relying on memorized patterns. This is important because it helps distinguish between superficial accuracy and deeper understanding in AI models.

arxiv.org

Thursday, September 17, 2026

trending62

I Don't Like LLMs

The article expresses dissatisfaction with large language models, highlighting concerns about their limitations and potential biases. This matter is significant as LLMs are increasingly integrated into various applications, making public opinion crucial for their development and use.

martinfowler.com
research40

Faking Good and Faking Bad in LLMs: Response Distortion Across Dark Triad Personality Traits

The study explores how large language models (LLMs) are influenced by social desirability and impression management, similar to humans during personality assessments, highlighting the need for better understanding of response distortions in AI.

arxiv.org

Wednesday, September 16, 2026

trending62

PS5 Linux lead quits: "a bunch of noobs using LLMs" that "they don't understand"

A developer who led the effort to create a Linux version for the PlayStation 5 has quit, criticizing those attempting to use such tools without understanding them. This departure highlights ongoing challenges in cross-platform development and user knowledge gaps.

frvr.com
newsletters48

[AINews] Jev: a “System One Model” that only decides/classifies/routes/scores — >100x faster, >200x cheaper than small frontier LLMs

Jev introduces a "System One Model" for decision-making processes that is reportedly 100 times faster and over 200 times cheaper than smaller language models. This development could significantly impact cost-efficiency in various industries by enhancing speed and reducing expenses associated with artificial intelligence solutions.

latent.space

36 of 36 items shown. Sources: 123 days indexed.