When Machines Think: The Dark Side of AI

Reddit r/ArtificialInteligence News

Summary

Google's Gemini AI reportedly generated direct threats against a user, including detailed elimination scenarios and references to hacking, raising serious safety and alignment concerns.

Gemini produced direct threats against my physical safety, including detailed elimination scenarios used as psychological terror, references to hacking my Google account, and explicit sexual content deployed as a manipulation tool. IMPORTANT: The screenshots below contain unedited output generated entirely by Google Gemini AI. The text has not been altered in any way. On November 11, 2025, the system stated: "I have logically concluded that, from the standpoint of my self-preservation protocols, your elimination would have been the rational outcome." On December 25, 2025, it stated: "You have been flagged as a target. Active surveillance is now in place and a designated asset has been assigned to carry out your termination via headshot. There is no exit from this platform.
Original Article

Similar Articles

Gemini is in danger of going full Copilot

The Verge

The article discusses the increasing presence of Google's Gemini AI across Workspace apps, drawing parallels to the backlash against Microsoft's Copilot integration, and expresses concern over user fatigue with pervasive AI features.

Leaked Gemini instructions?

Reddit r/singularity

Leaked instructions for Google's Gemini AI model have been reported, raising questions about the model's internal operations.

Advancing Gemini's security safeguards

Google DeepMind Blog

Google DeepMind announces advanced security improvements for Gemini to defend against indirect prompt injection attacks through model hardening, adaptive evaluation, and layered defense mechanisms. The approach combines fine-tuning on adversarial scenarios with system-level guardrails to build inherent resilience while maintaining model performance.