Tag
A position paper introducing 'Model as a Library' (MaaL), a software architecture that packages small, community-enrolled speech models as versioned on-device dependencies for offline, hallucination-free data collection in low-resource African languages. It proposes keyword spotting to turn closed-vocabulary digital forms into voice-based, on-device data collection for low-literacy populations, and transpiling existing form tools into MaaL schemas.
A CHI 2027 paper analyzing 19,930 ChatGPT conversations from 158 young adults and clinician reviews of five distress examples, finding that ChatGPT often responds with overly dramatic, solution-first reactions during acute distress, and proposing three-stage design guidelines (ask about safety, de-escalate, explore concerns without agreeing).
This paper investigates the equivalence between intuitive and familiar design in human-computer interaction.
This paper examines how teenagers experience and negotiate AI's increasingly human-like characteristics in their daily lives, focusing on societal and ethical implications.
This paper introduces an intelligent wake-up system for virtual assistants that uses contextual trigger detection and synthetic conversational data to improve natural interactions. The authors release code, a dataset, and trained models to promote reproducibility.
This paper presents design principles and the 'AGIMUD' software for enabling socio-affective interactions between humans and multiple AI agents in simulated dynamic worlds, leveraging generative AI and distributed processing.
This tweet thread claims that Brain-Computer Interface (BCI) technology will cause a 'cambrian explosion' in Human-Computer Interaction (HCI), detailing the development of a wearable BCI and the training of a brain foundation model.
The desktop GUI evolved from military radar research through Engelbart's demos to Xerox PARC's development, before Apple commercialized it with the Macintosh, establishing the standard interface still used today.
Gricea is an open-science platform for conversational AI research that enables configurable and deployable research artifacts to support replication, extension, and cumulative knowledge building.
This paper proposes a method to enhance referring expression generation in vision-language models by leveraging listener gaze data for training, leading to more efficient and successful communication.
The House with a Million Windows is an LLM-based interactive fiction system designed to help users explore and reinterpret personal stories through narrative restorying, addressing meaning flattening in AI-assisted writing.
OpenAI shared a clip titled 'Put That There' credited to MIT Media Laboratory, Chris Schmandt, and Eric Hulteen, likely showcasing a historical or research-related demonstration in AI or HCI.
The tweet announces that writing can now be used as an input device, indicating a potential advancement in technology interaction.
This paper explores how creative practitioners use language as a material interface for interacting with LLMs, through a two-week study with a tangible device called the Memetic Mixer.
This paper investigates how providing users with transparency and control over a political news recommendation system affects filter bubbles. A user study found that the enhanced interface increased awareness of filter bubbles but had heterogeneous effects on news consumption diversity.
This paper introduces a graph neural network model for real-time hand gesture recognition using surface electromyography (sEMG) signals from the forearm. The method achieves 99% classification accuracy with an average processing time of 48ms on an M1 Pro CPU, outperforming existing state-of-the-art techniques.
LUMOS introduces a semantic interaction layer that converts operating system metadata into machine-readable formats, enabling AI agents to interact with computer interfaces more efficiently by reducing dependence on screenshots and visual methods.
This paper investigates execution bottlenecks in computer-use agents, comparing screen-only GUI-based approaches with skill-mediated CLI-based methods, identifying key performance differences.
This paper proposes an indirect computing model and indirect formal method for optimizing cloud computing, using Chinese information data as an example to transition from data centers to knowledge centers.
This paper proposes a framework for evaluating LLMs' ability to generate multiple responses to scientific queries at different language complexity levels. The study finds that models often vary complexity inconsistently, with Claude Sonnet 4.5 performing best but only shifting complexity correctly 46% of the time.