Topic: safety

7 stories found

Yesterday

ai_labs75

Advancing responsible AI across Europe

OpenAI outlines its approach to ensuring AI safety and transparency to support responsible AI governance in Europe as the EU AI Act progresses. This is crucial for aligning with upcoming EU regulations and fostering trust in AI technologies.

openai.com↗

Thursday, July 30, 2026

research40

Steering Instruction Hierarchies at Inference Time

A recent study highlights that current large language models frequently disregard hierarchical instruction priorities, potentially compromising safety and reliability during deployment. This issue is crucial because proper adherence to these hierarchies ensures higher levels of control and predictability in how the models operate under conflicting directives.

arxiv.org↗

Wednesday, July 29, 2026

research40

MyoCardBench: A Real-World Data Benchmark for Evaluating Large Language Models in Clinically Authentic Cardiovascular Care Scenarios

MyoCardBench is a new benchmark for evaluating large language models in realistic cardiovascular care scenarios, addressing limitations of existing benchmarks which often focus on isolated tasks or knowledge rather than longitudinal, multimodal, and safety-critical clinical workflows.

arxiv.org↗

Monday, July 20, 2026

🌿 That's all for now. Come back tomorrow.

7 of 7 items shown. Sources: 73 days indexed.