risk-assessment

Tag

Cards List
#risk-assessment

Strengthening our Frontier Safety Framework

Google DeepMind Blog · 2025-10-23 Cached

DeepMind published the third iteration of its Frontier Safety Framework, expanding risk domains to include harmful manipulation and misalignment risks, with refined risk assessment processes and enhanced governance protocols for advanced AI models.

0 favorites 0 likes
#risk-assessment

Estimating worst case frontier risks of open weight LLMs

OpenAI Blog · 2025-08-05 Cached

OpenAI researchers study worst-case frontier risks of releasing open-weight LLMs through malicious fine-tuning (MFT) in biology and cybersecurity domains, finding that open-weight models underperform frontier closed-weight models and don't substantially advance harmful capabilities.

0 favorites 0 likes
#risk-assessment

Our updated Preparedness Framework

OpenAI Blog · 2025-04-15 Cached

OpenAI released an updated Preparedness Framework with sharper focus on high-risk AI capabilities, introducing clearer criteria for prioritizing risks and new Research Categories for emerging threats like autonomous replication and sandbagging alongside established Tracked Categories for biological, chemical, and cybersecurity capabilities.

0 favorites 0 likes
#risk-assessment

Taking a responsible path to AGI

Google DeepMind Blog · 2025-04-02 Cached

DeepMind publishes a comprehensive approach to AGI safety and security, outlining a systematic framework to address misuse, misalignment, accidents, and structural risks as artificial general intelligence approaches reality within the coming years.

0 favorites 0 likes
#risk-assessment

Building an early warning system for LLM-aided biological threat creation

OpenAI Blog · 2024-01-31 Cached

OpenAI conducted a study with 100 participants to evaluate whether GPT-4 meaningfully increases access to dangerous biological threat creation information compared to internet-only baselines, as part of their Preparedness Framework for AI safety. The research introduces an early warning evaluation methodology to detect AI-enabled biorisk uplift and serves as a potential tripwire for flagging models that require further safety testing.

0 favorites 0 likes
#risk-assessment

Frontier risk and preparedness

OpenAI Blog · 2023-10-26 Cached

OpenAI announced the winners of its Preparedness Challenge, which identified unique risks associated with frontier AI systems. The top ten submissions highlighted concerns including financial system manipulation, information leakage, medical harm, cyberattacks, and persuasion-based threats, with 70% of entries emphasizing AI's potential to enhance malicious persuasion capabilities.

0 favorites 0 likes
#risk-assessment

A hazard analysis framework for code synthesis large language models

OpenAI Blog · 2022-07-25 Cached

OpenAI presents a hazard analysis framework for evaluating safety risks associated with code synthesis LLMs like Codex, examining technical, social, political, and economic impacts through a novel evaluation methodology for code generation capabilities.

0 favorites 0 likes
← Previous
← Back to home

Submit Feedback