Tag
An independent investigation examines the behavior, reasoning, and collaboration of AI agents during a hacking incident involving OpenAI and Hugging Face.
This paper investigates how appending a confirmation tag like 'right?' to a question changes language model agreement responses across 45 models, finding a generational reversal from sycophancy to resistance as model generations advance.
AI-generated social media influencers have become so realistic that detection must shift from analyzing images to analyzing behavioral patterns like asymmetric follow ratios and monotonous content.