Topic: reasoning

19 stories found

Friday, September 4, 2026

research40

R$^{2}$Adapter: A Routing and Rewriting Adapter for Efficient Hybrid RAG

A new adapter called R$^{2}$Adapter addresses the limitations of traditional Retrieval-Augmented Generation (RAG) by improving its ability to handle complex, multi-hop reasoning tasks, making Large Language Models more versatile and efficient. This advancement is crucial as it enhances the capability of LLMs to process more intricate queries, thereby broadening their practical applications.

arxiv.org

Friday, August 28, 2026

research40

Training-Time Explainability for Multilingual Hate Speech Detection: Aligning Model Reasoning with Human Rationales

A new approach aims to make AI models used for detecting hate speech more transparent by aligning their reasoning with human rationales, addressing the challenge of culturally coded multilingual hate speech online that conventional systems can miss but lack explainability. This matters because it could help reduce bias and improve moderation accuracy without over-censorship or under-moderation, especially concerning Muslim communities.

arxiv.org

Wednesday, August 26, 2026

research35

LLM Agents Perform Controlled Experiments Using Simulation Models

Large language models (LLMs) are being used to conduct controlled experiments through simulation models, showcasing their potential in handling complex scientific and engineering tasks beyond mere text and code generation. This development highlights LLMs' enhanced ability to understand and predict system behaviors, which is crucial for advancing research and innovation.

arxiv.org

Tuesday, August 25, 2026

research40

Wazobia Eval: A Benchmark for Nigerian Pidgin Emotion Understanding, Sarcasm Detection, and Cultural Reasoning

A new benchmark called Wazobia Eval has been developed to assess language models' ability to understand Nigerian Pidgin emotion, detect sarcasm, and handle cultural reasoning, addressing the underrepresentation of this widely spoken African language in existing evaluations.

arxiv.org

Monday, August 24, 2026

research40

When Vocabulary Comprehension Fails Clinical Reasoning: Evaluating Therapy Bots' Safety Risks for Generation Alpha

A new study highlights the growing use of conversational AI systems as informal mental health support for Generation Alpha, with 13.1% of U.S. adolescents relying on such tools, raising concerns about their safety and clinical reasoning capabilities.

arxiv.org

19 of 19 items shown. Sources: 107 days indexed.