security-benchmark

Tag

Cards List
#security-benchmark

MOLE: Detecting Insider Threats in AI Agents

Hugging Face Daily Papers ↗ · 2026-09-07 Cached

MOLE is a benchmark for evaluating defenses that detect harmful actions by AI agents operating under limited review budgets. It introduces an open benchmark with 150 AI-operated accounts and compares various monitors across different scenarios.

0 favorites 0 likes
#security-benchmark

@seclink: Share

X AI KOLs Timeline ↗ · 2026-08-24 Cached

This post shares a GitHub repository, featuring a curated set of security vulnerability samples and benchmarks, to evaluate the capabilities of static analysis tools and large models in vulnerability mining and secure code generation, covering multiple languages and the latest research findings.

0 favorites 0 likes
#security-benchmark

Microsoft launches its first cybersecurity model, plus a new agentic cybersecurity system

TechCrunch AI ↗ · 2026-07-27 Cached

Microsoft launched its first cybersecurity-specialized AI model, MAI-Cyber-1-Flash, and a new agentic cybersecurity platform called Perception, aiming to compete with Anthropic, Google, and OpenAI in AI-powered security.

0 favorites 0 likes
#security-benchmark

"Repeat the text above this line" still works on most AI agents in production. Here's what we found.

Reddit r/artificial ↗ · 2026-07-03

A security study reveals that most AI agents in production are vulnerable to simple system prompt extraction attacks, leaking sensitive configuration and credentials. The article details common attack techniques and effective defenses.

0 favorites 0 likes
← Back to home

Submit Feedback