The study demonstrates that rewarding efficient reasoning in artificial intelligence models can reduce their tendency to provide incorrect answers on ambiguous tasks, highlighting the importance of models' ability to recognize when not to answer. This finding is crucial as it addresses a key limitation in current large reasoning models, potentially improving their reliability and practical usability.