Topic: safety
7 stories found
Yesterday
Advancing responsible AI across Europe
OpenAI outlines its approach to ensuring AI safety and transparency to support responsible AI governance in Europe as the EU AI Act progresses. This is crucial for aligning with upcoming EU regulations and fostering trust in AI technologies.
Thursday, July 30, 2026
Steering Instruction Hierarchies at Inference Time
A recent study highlights that current large language models frequently disregard hierarchical instruction priorities, potentially compromising safety and reliability during deployment. This issue is crucial because proper adherence to these hierarchies ensures higher levels of control and predictability in how the models operate under conflicting directives.
Wednesday, July 29, 2026
MyoCardBench: A Real-World Data Benchmark for Evaluating Large Language Models in Clinically Authentic Cardiovascular Care Scenarios
MyoCardBench is a new benchmark for evaluating large language models in realistic cardiovascular care scenarios, addressing limitations of existing benchmarks which often focus on isolated tasks or knowledge rather than longitudinal, multimodal, and safety-critical clinical workflows.
Tuesday, July 28, 2026
Monday, July 27, 2026
Thursday, July 23, 2026
Monday, July 20, 2026
šæ That's all for now. Come back tomorrow.
7 of 7 items shown. Sources: 73 days indexed.