user-study

Tag

Cards List
#user-study

How Does Empowering Users with Greater System Control Affect News Filter Bubbles?

arXiv cs.AI · 4d ago Cached

This paper investigates how providing users with transparency and control over a political news recommendation system affects filter bubbles. A user study found that the enhanced interface increased awareness of filter bubbles but had heterogeneous effects on news consumption diversity.

0 favorites 0 likes
#user-study

From Words to Widgets for Controllable LLM Generation

arXiv cs.CL · 2026-07-15 Cached

Malleable Prompting is a novel interactive technique that reifies natural language preferences into GUI widgets (sliders, toggles, dropdowns) for direct manipulation, with a decoding algorithm that modulates token probabilities based on widget values to enable precise control over LLM generation. A user study shows it outperforms natural language prompting in precision, controllability, and transparency.

0 favorites 0 likes
#user-study

Consensus vs. Dissent: Dynamic LLM Modeling of Subjective Preferences in Group Recommenders

arXiv cs.CL · 2026-07-14 Cached

This research fine-tunes LLMs on human survey data to serve as judgmental models for group recommender systems, dynamically selecting aggregation strategies to maximize satisfaction and consensus. A user study validates that the approach aligns with human fairness and satisfaction perceptions.

0 favorites 0 likes
#user-study

Anthropic studied 400K Claude Code sessions: domain knowledge mattered more than coding skill

Reddit r/ArtificialInteligence · 2026-06-17

Anthropic analyzed 400K Claude Code sessions and found that domain expertise is a stronger predictor of success than coding skill, with experts achieving 28-33% verified success versus 15% for novices. The study highlights that understanding the problem matters more than coding ability.

0 favorites 0 likes
#user-study

Security and Privacy Prompts in the Wild: What Users Ask LLMs and How LLMs Respond

arXiv cs.CL · 2026-06-17 Cached

This paper analyzes real-world user queries about digital security and privacy asked to LLMs, categorizing them into nine topics and evaluating response quality and consistency across commercial and open-weight models.

0 favorites 0 likes
#user-study

Nonslop: A Gamified Experiment in Human-AI Collaborative Writing

arXiv cs.AI · 2026-06-11 Cached

This paper presents a gamified experiment where participants write responses with AI suggestions disincentivized, analyzing when humans adopt AI assistance versus maintaining creative autonomy.

0 favorites 0 likes
#user-study

Characterizing initial human-AI proof formalization workflows

arXiv cs.AI · 2026-06-04 Cached

Researchers from Oxford, Cambridge, MIT, CMU and other institutions conduct a mixed-methods study examining how people integrate AI tools into mathematical proof formalization workflows, finding that participants generally achieve higher formalization accuracy with AI assistance while preferring to maintain high-level human control over the proof discovery process.

0 favorites 0 likes
#user-study

Effects of Varying LLM Access on Essay Writing Behavior

arXiv cs.CL · 2026-06-02 Cached

A pilot study with 24 college students examines how varying levels of LLM access (none, limited, unlimited) affect essay writing quality, behavior, and perceived authorship, finding that constrained access preserves authorship confidence while unlimited access reduces creative expression and ownership.

0 favorites 0 likes
#user-study

PrivacyAkinator: Articulating Key Privacy Design Decisions by Answering LLM-Generated Multiple-choice Questions

arXiv cs.AI · 2026-05-22 Cached

This paper presents PrivacyAkinator, an interactive tool that helps novice developers articulate privacy design decisions via LLM-generated multiple-choice questions, achieving 47% more key decisions in 73% less time compared to NIST's PRAM methodology.

0 favorites 0 likes
#user-study

Evaluating multimodal emotion recognition in proactive conversational agents: A user study

arXiv cs.AI · 2026-05-22 Cached

This paper presents a multimodal emotion recognition module for proactive conversational agents, using facial recognition and linguistic analysis. A user study with 20 participants reveals a 'poker face' effect where visual cues are unreliable, while linguistic analysis proves more accurate; the study also shows agents can elicit emotions through conversational adaptation.

0 favorites 0 likes
#user-study

"I didn't Make the Micro Decisions": Measuring, Inducing, and Exposing Goal-Level AI Contributions in Collaboration

Hugging Face Daily Papers · 2026-05-20 Cached

Introduces CoTrace, a framework for goal-level attribution in human-AI collaboration, which analyzes how large language models shape goals by contributing concrete requirements and indirect influences in dialogue turns.

0 favorites 0 likes
#user-study

Beyond Autonomy: The Power of an Agent That Knows Its Limits

Reddit r/AI_Agents · 2026-05-08

The COWCORPUS project, a study of 4,200 human-AI interactions, found that agents predicting their own failures and intervention moments are more useful than those simply trying to avoid errors. Researchers identified four stable trust patterns in human-AI collaboration and developed the Perfect Timing Score (PTS) to measure intervention prediction accuracy.

0 favorites 0 likes
#user-study

Apr 30, 2026Societal ImpactsHow people ask Claude for personal guidance

Anthropic Research · 2026-05-08 Cached

Anthropic presents research on how users seek personal guidance from Claude, highlighting findings on sycophancy rates across domains. The study informed the training of Claude Opus 4.7 and Mythos Preview to better protect user wellbeing.

0 favorites 0 likes
← Back to home

Submit Feedback