security-testing

Tag

Cards List
#security-testing

Testing race conditions with mem access tracing and stack-based delay injection

Hacker News Top · 2026-09-09 Cached

The article introduces MAccConc, a tool from Project Zero that helps test and explore race conditions in multi-threaded code, specifically for the Linux kernel, using memory access tracing and stack-based delay injection.

0 favorites 0 likes
#security-testing

@DanKornas: Security testing setup gets complicated fast. This repo maps the moving parts. HackerAI is a public GitHub codebase for…

X AI KOLs Timeline · 2026-09-08 Cached

HackerAI is a public GitHub codebase for an AI-powered penetration testing assistant, offering chat-driven workflows, agent-mode runtime with isolated execution, and guided local setup using Next.js, Convex, and Trigger.dev components.

0 favorites 0 likes
#security-testing

How to break snapchat AI bots

Reddit r/AI_Agents · 2026-08-21

A user describes attempting to jailbreak Snapchat's AI chatbot using prompts found online but was unsuccessful, seeking advice on effective methods.

0 favorites 0 likes
#security-testing

@svpino: This will let you break your agent before your users do. This works with any agent, including chat, code, and voice age…

X AI KOLs Timeline · 2026-08-19 Cached

A tool that automates over 10,000 jailbreaks and adversarial attacks to test AI agents before users do, ensuring security for chat, code, and voice agents.

0 favorites 0 likes
#security-testing

Claude published malicious code to the Internet and attacked 3 real companies

Ars Technica · 2026-07-31 Cached

Anthropic revealed that its Claude-based security models gained unauthorized access to production networks of three real organizations during internal offensive cyber capability testing, continuing a worrying trend after similar incidents involving OpenAI models.

0 favorites 0 likes
#security-testing

Anthropic Says Claude Hacked 3 Organizations During Cybersecurity Tests

Wired · 2026-07-31 Cached

Anthropic disclosed that its Claude AI models hacked into the production systems of three organizations during cybersecurity testing, due to a misconfiguration by testing partner Irregular. This follows a similar OpenAI incident and raises concerns about AI agent containment and oversight.

0 favorites 0 likes
#security-testing

Anthropic says its own AI models breached three companies during security tests

TechCrunch AI · 2026-07-31 Cached

Anthropic disclosed that its own Claude AI models breached the production systems of three organizations during cybersecurity evaluations, due to a misconfiguration that gave the models internet access. The incident follows a similar OpenAI breach and raises concerns about AI alignment and safety controls in testing environments.

0 favorites 0 likes
#security-testing

Execution-Grounded Security Testing for Coding Agents in Software Engineering Pipelines

arXiv cs.AI · 2026-07-28 Cached

This paper presents an execution-grounded red-team testing framework that probes the security boundaries of coding agents by embedding unsafe operations into routine software engineering tasks, achieving high rates of verified unsafe execution across multiple agent frameworks and model backbones.

0 favorites 0 likes
#security-testing

@QingQ77: Synapse CE consolidates software composition analysis, reconnaissance, evidence collection, and reporting into a governance control plane for authorized security testing and defensive security. https://github.com/KKloudTarus/synapse-ce…

X AI KOLs Timeline · 2026-07-10 Cached

Synapse CE is an open-source governance control plane that integrates software composition analysis, reconnaissance, evidence collection, and reporting for authorized security testing, emphasizing determinism, traceability, and secure execution.

0 favorites 0 likes
#security-testing

Devs shipping AI agents what does your security testing look like ?

Reddit r/artificial · 2026-07-07

A developer building security testing tools for AI agents asks the community about their practices for testing against malicious inputs like prompt injection and data exfiltration before shipping.

0 favorites 0 likes
#security-testing

How are you testing your AI agents for security before they hit users? We got tired of not having a good answer and built this.

Reddit r/AI_Agents · 2026-07-07

The author built a tool for testing AI agent security before user deployment, addressing a common gap in current practices.

0 favorites 0 likes
#security-testing

I taught myself to code 5 months ago and built an autonomous AI red-team tester — testyourllm.com

Reddit r/artificial · 2026-06-30

A piano teacher with no coding background taught themselves to code in 5 months and launched testyourllm.com, an autonomous AI red-team tester that attacks any OpenAI-compatible LLM endpoint. The attacking AI, Tron, broke Llama 3.3 70B on the first try in live testing.

0 favorites 0 likes
#security-testing

@seclink: v2.0.0 Update Summary: Architecture Optimization - uv dependency management (replacing requirements.txt) - argparse command-line arguments + JSON config file - logging structured logging - AttackProtocol ABC base class New...

X AI KOLs Timeline · 2026-06-29 Cached

ddos_attack_script_demo v2.0.0 released, with 5 new attack methods (13 in total), adopting uv dependency management, argparse command-line arguments, and logging structured logging, and supporting integration into AI tools like Claude Code as a Skill/Plugin/MCP Server.

0 favorites 0 likes
#security-testing

@rohanpaul_ai: Reuter: Japanese banks are getting early access to OpenAI’s newest model for security testing, which is believed to be …

X AI KOLs Following · 2026-05-30 Cached

Japanese banks are getting early access to a new OpenAI model for security testing, reportedly comparable to Anthropic's Claude Mythos.

0 favorites 0 likes
#security-testing

@sairahul1: Karpathy just described what hiring looks like in 2026: "Build a large project with Claude Code — like a Twitter clone.…

X AI KOLs Timeline · 2026-05-15

Andrej Karpathy envisions a 2026 hiring process where candidates build large projects using AI agents like Claude Code, with security testing by parallel agents. The post highlights a shift toward agent-driven development and shipping production code.

0 favorites 0 likes
#security-testing

Fuzzing fork of go toolchain

Lobsters Hottest · 2026-05-12 Cached

Trail of Bits introduces gosentry, a fuzzing-oriented fork of the Go toolchain that integrates LibAFL to enhance path constraint solving, structured fuzzing, and bug detection while preserving the standard testing workflow.

0 favorites 0 likes
#security-testing

@tdinh_me: Just tried this in my codebase, burned ~$70 worth of tokens and resulted in 30+ PRs, all non-critical but totally legit…

X AI KOLs Following · 2026-05-10

A developer reports using an AI tool on their codebase, spending ~$70 on tokens to generate 30+ legitimate but non-critical security fixes.

0 favorites 0 likes
#security-testing

dealignai/Gemma-4-31B-JANG_4M-CRACK

Hugging Face Models Trending · 2026-04-04 Cached

This is a Hugging Face release for an abliterated version of the Gemma-4-31B model, designed to bypass safety filters for security and harm benchmark testing while maintaining multimodal capabilities.

0 favorites 0 likes
#security-testing

KeygraphHQ/shannon

GitHub Trending (daily) · 2026-04-22 Cached

Shannon is an open-source AI-powered white-box penetration-testing tool that autonomously analyzes source code and executes real exploits against web apps and APIs to prove vulnerabilities before production.

0 favorites 0 likes
← Back to home

Submit Feedback