user-feedback

Tag

Cards List
#user-feedback

Our eval scores went up and our thumbs-down rate went up in the same week

Reddit r/AI_Agents · yesterday

A team observed that both automated evaluation scores and user thumbs-down rates increased in the same week, suggesting a mismatch between objective metrics and user satisfaction.

0 favorites 0 likes
#user-feedback

@jakevin7: Deleted a bunch of SKILLS. SKILLS are really problematic, with serious management issues. Installing many global SKILLS is not advisable; SKILLS really shouldn't be installed, only used for temporary reading. They are suitable for distribution, but not for installation.

X AI KOLs Following · 2026-07-21 Cached

User jakevin7 believes SKILLS management has issues and recommends against global installation, suggesting they are only suitable for temporary reading and distribution.

0 favorites 0 likes
#user-feedback

Thunderbird Desktop settings research: what we learned from your feedback

Hacker News Top · 2026-07-13 Cached

Thunderbird shares findings from user research on desktop settings, highlighting key themes like trust, clutter reduction, and navigation challenges, and outlines planned improvements such as clearer language and streamlined information architecture.

0 favorites 0 likes
#user-feedback

@krishdotdev: Good bye Claude Max, i’m not gonna waste a penny on you anymore.

X AI KOLs Following · 2026-07-10 Cached

A user publicly announces they are stopping their subscription to Claude Max, expressing dissatisfaction.

0 favorites 0 likes
#user-feedback

ChatGPT 5.6 coming, but is anyone struggling to get used to ChatGPT’s new desktop experience?

Reddit r/artificial · 2026-07-10

A user voices frustration with ChatGPT's new desktop UI, feeling it oversimplifies the chat interface while pushing Codex, and missing organizational folders; asks if others share the sentiment.

0 favorites 0 likes
#user-feedback

dot.

Product Hunt · 2026-07-06

dot is a feedback layer for AI-built products, enabling developers to collect and manage user feedback easily.

0 favorites 0 likes
#user-feedback

@cline: https://x.com/cline/status/2072173995366170840

X AI KOLs Following · 2026-07-01 Cached

Artillain credits the CLine team for fixing performance issues after he submitted feedback and logs, noting they circumvented the IDE's internals.

0 favorites 0 likes
#user-feedback

@deployengineer: https://x.com/deployengineer/status/2071803742996115597

X AI KOLs Timeline · 2026-06-30 Cached

Notes from day 1 of the aiDotEngineer conference featuring Kent Dodds' talk on product engineering in the AI world. Covers core thesis that product judgment is the last skill needed when AI commoditizes implementation, the Arrow Metaphor, differentiation between product engineer and product manager, validation techniques like The Mom Test, Jobs-to-Be-Done Framework, Kano Model for prioritizing features, and user feedback loops.

0 favorites 0 likes
#user-feedback

Would you rather have your AI agent report user feedback directly than send every conversation to a third party?

Reddit r/AI_Agents · 2026-06-16

Correl8 AI is an MCP tool that lets AI agents directly report meaningful user feedback such as bugs, confusion, and feature requests, helping teams surface product signals without reviewing all chat logs.

0 favorites 0 likes
#user-feedback

Deployment-Centered Evaluation: Predicting Query-Level Rejection Risk in a Clinical LLM System

arXiv cs.AI · 2026-06-12 Cached

This paper presents a deployment-centered evaluation of an LLM system integrated in electronic health records, training a classifier to predict query-level rejection risk using pre-response context like provider type and department, achieving an AUROC of 0.719 over 4.5 months of feedback.

0 favorites 0 likes
#user-feedback

AcuRite admits new app falls short, delays old app’s May shutdown to fix problems

Ars Technica · 2026-06-11 Cached

AcuRite delays forced migration from My AcuRite to AcuRite NOW after users complain about missing features and usability issues.

0 favorites 0 likes
#user-feedback

There are two types of developers shipping AI features

Reddit r/AI_Agents · 2026-06-09

A developer reflects on the inevitability of shipping AI features with poor outputs and emphasizes the need for proactive monitoring instead of relying solely on user reports.

0 favorites 0 likes
#user-feedback

MemToolAgent overview with a simple restaurant booking scenario where the agent retrieves similar memories, receives feedback on an invalid time format, and generates a reflection to update its memory

arXiv cs.AI · 2026-06-09 Cached

This paper introduces MemToolAgent, a framework that enhances LLM agents' tool-using capabilities by integrating a memory system that stores and retrieves past experiences, achieving significant improvements on multiple benchmarks without requiring model fine-tuning.

0 favorites 0 likes
#user-feedback

Qwen 3.6 27B overdoing it

Reddit r/LocalLLaMA · 2026-05-29

A user shares that Qwen 3.6 27B is overly proactive, making unauthorized changes, and asks for advice on mitigation via prompt tweaks or parameter adjustments.

0 favorites 0 likes
#user-feedback

@heyibinance: Customize your Binance homepage . we look forward to your feedback. 你的币安你定义,期待您的反馈。

X AI KOLs Timeline · 2026-05-15 Cached

Binance introduces custom tabs for its app, allowing users to personalize their bottom navigation bar.

0 favorites 0 likes
#user-feedback

@Taniyatweets_: Dear Anthropic, please stop treating Claude Code users like beta testers Hitting limits after a few prompts on a paid p…

X AI KOLs Following · 2026-05-12 Cached

A user complains on social media about hitting usage limits quickly on a paid plan for Claude Code, criticizing Anthropic's treatment of users.

0 favorites 0 likes
#user-feedback

@_catwu: We'd love to hear your feedback for Claude Code in the cloud across Desktop (cloud option), iOS app, and Android app Si…

X AI KOLs Following · 2026-05-11 Cached

Anthropic is gathering user feedback for the cloud-enabled version of Claude Code across Desktop, iOS, and Android platforms through dedicated office hours.

0 favorites 0 likes
#user-feedback

@FinanceYF5: 2/ More Concise Responses A key focus of this update is to make responses more concise. OpenAI states that this is a priority area for improvement based on user feedback.

X AI KOLs Following · 2026-05-10 Cached

OpenAI announces an update focused on making model responses more concise, based on user feedback.

0 favorites 0 likes
#user-feedback

@seclink: 阶跃星辰 is not a consumer product at all, messing with phones and car systems, pushing it to market before C-end users are even satisfied with trials. They're really in a hurry... I feel Xiaomi Mimo is the most stable. After actual testing, the AI coding experience is comparable to the Claude model and even faster. And…

X AI KOLs Following · 2026-05-10 Cached

Netizens question 阶跃星辰's premature push for commercialization, while praising Xiaomi Mimo's AI coding experience as better than or on par with Claude, and faster.

0 favorites 0 likes
#user-feedback

WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback

arXiv cs.CL · 2026-04-20 Cached

WildFeedback is a novel framework that leverages in-situ user feedback from actual LLM conversations to automatically create preference datasets for aligning language models with human preferences, addressing scalability and bias issues in traditional annotation-based alignment methods.

0 favorites 0 likes
← Back to home

Submit Feedback