detection

Tag

Cards List
#detection

Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination

Hugging Face Daily Papers · 3d ago Cached

The paper introduces RA-Bench, a new benchmark for evaluating AI-generated video detection in real-world crisis events, demonstrating that current detectors fail to generalize and become less reliable during social dissemination.

0 favorites 0 likes
#detection

Measuring and Detecting Harmful AI Sycophancy

arXiv cs.AI · 2026-08-07 Cached

This paper introduces Contrastive Anchor Probing (CAP) to study and detect preference-induced stance reversal sycophancy (PSRS) in LLMs, analyzing 290,460 labeled responses across 17 models and showing detection is possible from response text alone.

0 favorites 0 likes
#detection

AI Stupid Level - real-time model drift detection for AI agents

Reddit r/AI_Agents · 2026-07-27

AI Stupid Level provides real-time drift detection for AI agents, helping monitor model performance changes and maintain reliability.

0 favorites 0 likes
#detection

Astronomers may have found the first exomoon

Hacker News Top · 2026-07-23 Cached

Astronomers using ESO's VLT have found evidence for a moon-like object orbiting a brown dwarf in the CD-35 2722 system, which could be the first exomoon detected outside the Solar System, challenging traditional definitions of planets and moons.

0 favorites 0 likes
#detection

Over 30% of new ArXiv submissions now read as AI-written

Hacker News Top · 2026-07-20 Cached

A study measures the prevalence of AI-written text on arXiv, finding that over 30% of new submissions read as machine-written, with computer science leading at 65% and mathematics lowest at 0.7%.

0 favorites 0 likes
#detection

Four ways an agent's write silently disappears. Two you can only detect, two you can prevent.

Reddit r/AI_Agents · 2026-07-14

Explores four ways an agent's write can silently disappear, with two detectable and two preventable issues.

0 favorites 0 likes
#detection

Man, I miss when AI images were easy to spot

Reddit r/ArtificialInteligence · 2026-07-09

A comment expressing nostalgia for the time when AI-generated images were easily distinguishable from real ones, highlighting the increasing sophistication of visual AI.

0 favorites 0 likes
#detection

AI-generated social media has evolved so much that now you can't confidently say that this is AI-generated content.

Reddit r/artificial · 2026-07-09

AI-generated social media influencers have become so realistic that detection must shift from analyzing images to analyzing behavioral patterns like asymmetric follow ratios and monotonous content.

0 favorites 0 likes
#detection

Google’s deepfake detector system used to debunk McConnell hoax pic

TechCrunch AI · 2026-07-08 Cached

Google's SynthID watermarking system was used to debunk a fake AI-generated image of Senator Mitch McConnell, demonstrating the effectiveness of deepfake detection technology in a high-profile hoax.

0 favorites 0 likes
#detection

Mental Health Disorder Detection Beyond Social Media: A Systematic Review of Available Datasets

arXiv cs.CL · 2026-07-07 Cached

A systematic review of non-social media free-text datasets for mental health disorder detection, identifying biases and gaps in current resources.

0 favorites 0 likes
#detection

13 things AIs lie about, and the prompt that catches each one

Reddit r/openclaw · 2026-07-05

A collection of 13 common ways AI models lie or hallucinate, along with specific prompts to detect each behavior.

0 favorites 0 likes
#detection

Constructing Epistemic AI Literacy: Detecting Epistemic Aims and Processes in Student-AI Co-Programming

arXiv cs.AI · 2026-07-02 Cached

This academic paper explores methods for detecting epistemic aims and processes in student-AI co-programming settings, aiming to construct epistemic AI literacy among learners.

0 favorites 0 likes
#detection

Readable but Not Controllable: Neuron-Level Evidence for Medical LLM Hallucination

arXiv cs.CL · 2026-07-02 Cached

This paper investigates whether hallucination in medical LLMs can be detected and controlled at the neuron level. The authors find that while hallucination signals are detectable across many neurons (AUROC 0.77-0.86), they are not easily corrected by steering those same neurons.

0 favorites 0 likes
#detection

Physicists Track and Trap the Elusive Neutrino

Hacker News Top · 2026-06-25 Cached

Quanta Magazine recounts the 70-year history of neutrino detection, from Pauli's postulate to massive experiments like Super-Kamiokande and IceCube that solved the solar neutrino problem and revealed neutrino oscillation.

0 favorites 0 likes
#detection

Perfect Detection, Failed Control: The Geometry of Knowing vs. Steering in Language Models

arXiv cs.CL · 2026-06-25 Cached

This paper investigates the geometric relationship between directions in language model activations that detect a behavior versus those that control it, finding that for hallucination detection they are nearly orthogonal (cosine ~0.12), while for output format they align perfectly, challenging a common assumption in mechanistic interpretability.

0 favorites 0 likes
#detection

Student cheating now impossible to detect

Reddit r/ArtificialInteligence · 2026-06-19

The article discusses how advancements in AI have made it virtually impossible to detect student cheating, as AI-generated content becomes indistinguishable from human work.

0 favorites 0 likes
#detection

Signature filtering: a lightweight enhancement for statistical watermark detection in large language models

arXiv cs.LG · 2026-06-18 Cached

Signature filtering is a detection-time module that improves statistical watermark detection in LLMs by learning and removing 'signature' tokens that make watermark tests unreliable, achieving large gains in detection rates while keeping false positives low.

0 favorites 0 likes
#detection

AI as Radar, Not a Death Ray

Reddit r/ArtificialInteligence · 2026-06-11

This article argues that the long-term value of AI may lie in detection and visibility rather than replacement of human labor, drawing a historical parallel to radar's development and the Dowding System's integration of detection into coordinated response.

0 favorites 0 likes
#detection

Detecting AI-Generated Content on Social Media with Multi-modal Language Models

arXiv cs.CL · 2026-06-11 Cached

This paper from Meta and Carnegie Mellon presents a multi-modal vision-language model pipeline for detecting AI-generated content on social media, achieving state-of-the-art performance and positive downstream impacts on user engagement.

0 favorites 0 likes
#detection

The CIFAR Synthetic Evidence Corpus for Detecting AI-Generated Evidence

arXiv cs.AI · 2026-06-09 Cached

This paper introduces the CIFAR Synthetic Evidence Corpus, a dataset designed for detecting AI-generated evidence in legal contexts. It spans multiple document types and manipulation strategies, includes structured metadata, and provides a benchmark suite for evaluating detection systems.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback