Topic: bias

6 stories found

Monday, August 31, 2026

Friday, August 28, 2026

research40

Training-Time Explainability for Multilingual Hate Speech Detection: Aligning Model Reasoning with Human Rationales

A new approach aims to make AI models used for detecting hate speech more transparent by aligning their reasoning with human rationales, addressing the challenge of culturally coded multilingual hate speech online that conventional systems can miss but lack explainability. This matters because it could help reduce bias and improve moderation accuracy without over-censorship or under-moderation, especially concerning Muslim communities.

arxiv.org

Wednesday, August 26, 2026

research35

A survey detection channel overrides the pixels in an astronomical foundation model, and biases tomographic mean redshifts

A study finds that foundation models in astronomy, trained on survey data including incomplete catalogues, inherit biases related to pixel-level incompleteness, potentially skewing mean redshift measurements. This matters because it highlights systemic issues in how astronomical data is processed and analyzed, impacting the accuracy of cosmological observations.

arxiv.org

Tuesday, August 25, 2026

research40

Mitigating Bias in Large Vision-Language Models via Counterfactual Ensemble Decoding

Large Vision-Language Models (LVLMs), while effective in various tasks, can exhibit biased behavior due to social biases in their training data. Researchers propose a method called Counterfactual Ensemble Decoding to mitigate these biases, highlighting the importance of addressing fairness in AI systems.

arxiv.org

Monday, August 24, 2026

research40

Who Do Language Models Think Is Competent? A Mechanistic Analysis of Occupational Bias

The study examines whether language models that pass behavioral bias tests still hold internal biases related to occupational competence, finding that they do retain such biases internally. This matters because it highlights the need for more comprehensive evaluation methods beyond surface-level behavior to ensure unbiased AI systems.

arxiv.org

🌿 That's all for now. Come back tomorrow.

6 of 6 items shown. Sources: 107 days indexed.