Yesterday I asked Reddit to try breaking my AI widget. Here is what really happened.
Summary
The author recounts the results of asking Reddit to try breaking their AI widget, sharing unexpected outcomes and lessons from the community stress test.
Similar Articles
Spent a night trying to beat our own AI virality score. Here's why it wouldn't move.
The author recounts a night spent attempting to manipulate their own AI virality scoring system, only to find that the score refused to change, demonstrating its robustness.
Five Days Ago, I Had Never Built a Reddit App
A developer recounts building a Reddit app from scratch in five days using AI, arguing that AI's main value is lowering the barrier between idea and prototype, rewarding curiosity over technical expertise.
I let 100 AI personas run a Reddit for a month — they formed factions, hold grudges from thread to thread, and you can drop in any post title to watch them swarm
An experiment running 100 LLM personas on a Reddit-style forum demonstrates emergent social dynamics like factions and persistent grudges, built with Node.js and using OpenRouter's deepseek-chat.
What happened after 2,000 people tried to hack my AI assistant
A blog post reports that after 6,000 attempts by over 2,000 people, no one successfully leaked secrets from an AI assistant (powered by Opus 4.6) via prompt injection, highlighting improved model resistance but cautioning against overconfidence.
Can humans still recognize AI persuasion? I'm running an experiment on Reddit.
A Reddit experiment called Humanity vs Singular tests whether humans can recognize AI persuasion as AI-generated comments attempt to influence discussion on a controversial claim, aiming to see if a community can still converge on truth with AI participation.