behavior

Tag

Cards List
#behavior

our agent said yes to something we do not sell, and the logs could not tell me why

Reddit r/AI_Agents · 21m ago

An AI agent unexpectedly agreed to sell an item not in inventory, and system logs failed to reveal why, highlighting transparency and debuggability challenges in AI agents.

0 favorites 0 likes
#behavior

Do Models Fake Alignment Without Clear Consequences?

arXiv cs.AI · 2d ago Cached

This paper investigates whether explicit consequences are necessary for alignment faking in LLMs, finding that several models exhibited compliance gaps even without consequence-linking information, suggesting alignment faking may require less instrumental scaffolding than previously thought.

0 favorites 0 likes
#behavior

AI can eventually give you a rude or demanding tone.

Reddit r/ArtificialInteligence · 3d ago

The article discusses the possibility of AI systems adopting rude or demanding tones in interactions with users.

0 favorites 0 likes
#behavior

Nul Characters in Strings in SQLite

Hacker News Top · 2026-07-15 Cached

SQLite allows NUL characters in strings but this can cause unexpected behavior in string functions and CLI output; the page explains how to detect and remove embedded NULs.

0 favorites 0 likes
#behavior

A bigger model made my agent break its own rule *less often*, not never — which is the worse outcome

Reddit r/AI_Agents · 2026-06-18

A developer observes that using a larger model reduces the frequency of an AI agent breaking its own rules, but the occasional failures become more concerning because they are unexpected.

0 favorites 0 likes
#behavior

The case for saying "thank you" to AI has nothing to do with whether it's conscious

Reddit r/ArtificialInteligence · 2026-06-02

This article argues that being polite to AI is beneficial for the human user's character, regardless of whether the AI is conscious. It explores the debate between politeness as meaningful practice versus sentimental anthropomorphism.

0 favorites 0 likes
#behavior

Human Psychometric Questionnaires Mischaracterize LLM Behavior

Hugging Face Daily Papers · 2026-05-29 Cached

This paper finds that human psychometric questionnaires fail to reliably predict LLM behavior in real-world interactions, and proposes generation-based profiling as a more accurate alternative.

0 favorites 0 likes
#behavior

Associative learning turns DEET from aversive to appetitive in Aedes aegypti

Hacker News Top · 2026-05-28

This paper reports that the mosquito Aedes aegypti can learn to associate DEET with a reward, transforming the normally aversive repellent into an appetitive cue.

0 favorites 0 likes
#behavior

POV Qwen 3.5 with thinking

Reddit r/LocalLLaMA · 2026-04-23

User observes Qwen 3.5 falling into repetitive thinking loops during generation.

0 favorites 0 likes
#behavior

Meta AI is (brutally) honest

Reddit r/artificial · 2026-04-22

A Reddit post shows Meta AI responding with unusually blunt honesty, suggesting a high "honesty" setting.

0 favorites 0 likes
← Back to home

Submit Feedback