Search arXiv⌕ Search

arXiv · 2610.08067

Frontstage Mediation Work: Invisible Work Bridging Gaps Between AI Decisions and User Expectations

Abstract

Automated service systems increasingly generate algorithmic operational decisions that shape how services are delivered. However, these decisions reach end-users only through frontline workers who carry them out in real-world settings. During this process, automated decisions can diverge from user expectations, surfacing as friction at the service encounter. We propose Frontstage Mediation Work as a preliminary analytic lens for examining the often invisible labor through which frontline workers anticipate and manage such misalignments between algorithmic decisions and user expectations. Drawing on a qualitative case study of an On-Demand Ride-Pooling service, we identify four recurring practices through which drivers sustain the service encounter when frictions arise. Such labor remains absorbed into routine operations, leaving no trace in performance metrics, system logs, or formal job descriptions. This paper contributes to worker-centered HCI scholarship by illustrating how automated services shift onto frontline workers the responsibility of managing the interactional consequences of system-level decisions.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Yongjae Sohn, Daehyun Kwak, Jiyeon Amy Seo, Hyungjun Cho, Seongah Youn, Youn-kyung Lim. 2026-10-06. Frontstage Mediation Work: Invisible Work Bridging Gaps Between AI Decisions and User Expectations. https://doi.org/10.1145/3785651.3831515

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

MATE: Diagnosing Empathy Calibration Failures in Multi-Turn Human-LLM Interaction

Large language models are increasingly used in emotionally consequential interactions. Response-level evaluation, however, struggles to diagnose how empathy fails across turns. We introduce MATE (Multi-turn Assessment of calibraTed Empathy), a framework that evaluates empathy as a multi-turn, perception-centered process. Across two controlled studies (N=82), baseline responses exhibit disclosure-insensitive miscalibration: generic validation and formulaic tone read as hollow across turns. A prompt-level self-critique condition makes these patterns less prominent but coincides with a different failure, agreement-skewed miscalibration, where affirmation becomes insufficiently contingent on disclosure context under high persona alignment (62.5% of fit participants in Study 2). Response-level analysis is consistent with a tone-agreement dissociation, as the condition shifts tone toward naturalness while agreement density also increases (+98% under fit vs. +25% under unfit). These findings reframe empathic behavior as relational calibration-consistency among disclosure depth, persona, and response behavior-rather than maximization. Response behavior itself spans tone, agreement, and contextual specificity, sub-dimensions that can drift independently.

cs.HC↗

CHOMP: Multimodal Chewing Side Detection with Earphones

Chewing-side preference (CSP) is a risk factor for temporomandibular disorders (TMDs) and a behavioral manifestation. Although TMDs affect roughly one-third of the global population, assessment relies on clinical examinations and self-reports, providing limited insight into everyday jaw function. We present CHOMP, the first earphone-based chewing-side detection system for continuous CSP monitoring. Using OpenEarable 2.0, we collected multimodal data from 20 participants with microphones, a bone-conduction microphone, IMU, PPG, and a pressure sensor across diverse foods, activities, and acoustic-interference conditions. CHOMP models paired-ear temporal feature sequences using modality-specific bidirectional GRUs, multimodal fusion, and prototype-based classification with an optional short user adaptation. Microphones achieve the strongest single-sensor performance, with median macro F1 scores of 97.2% under leave-one-food-out (LOFO) and 95.7% under leave-one-subject-out (LOSO) evaluation after user adaptation. Multimodal fusion reaches 98.0% under LOFO and 97.3% under adapted LOSO. We demonstrate CHOMP's performance under three acoustic-interference conditions and within a cafeteria setting. Our results establish earphones as a practical platform for everyday CSP monitoring and jaw-function assessment.

cs.HC↗

Tutor, Not Solver: Designing a Guardrailed AI Assistant for Learning in Higher Education: A Design Case of PeteChat

Generative artificial intelligence (AI) tutors hold significant promise for higher education, yet designing systems that scaffold learning without undermining academic integrity remains an open design challenge. This paper presents PeteChat, a course-aligned AI tutor developed and piloted at a large U.S. research university and documented through the lens of design-based research (DBR). Drawing on formative expert evaluation with teaching assistants and user-experience and developer stakeholders, we report eight transferable design principles for assessment-aware AI tutors, ranging from homework guardrails and debugging scaffolds to self-regulated learning support and instructor-facing customization tools. The system is built on a locally hosted large language model from the Llama-3 family, enhanced with retrieval-augmented generation (RAG) grounded in course-specific materials. Rather than reporting controlled experimental outcomes, this design case foregrounds the situated design reasoning, iterative refinement, and principled decision-making that shaped PeteChat across four development phases. The principles and methodological approach offer actionable guidance for institutions seeking to deploy responsible, integrity-preserving AI tutors at scale.

cs.HC↗