EdTech Discovery
Argus

Named after the hundred-eyed watchman of Greek myth, Argus watches the education landscape: spotting new opportunities, pressure-testing the ventures we're building, and tracing every read back to the real-world signals behind it.

Updated Aug 31, 2026 · 36 ideas · 18349 signals
Admin mode. Curation controls visible. Keep this URL (with token) private.

Signals

The evidence library: the raw signals the pipeline is watching across the education ecosystem. Every idea is built from these.

technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Floor, Ceiling, and the Fusion Gap: How Much of Crowd Reading Attention Can Machines Predict?

arXiv:2608.01704v1 Announce Type: cross Abstract: A benchmark score means nothing without knowing what a trivial method achieves and what the best possible method could achieve. We construct both bounds for a task with a rare kind of ground truth: predicting which sentences a crowd of readers -- highlighting for their own purposes, unpaid, uninstructed, and blind to each other -- marked in 120 web documents. The floor is naive truncation (lead); the ceiling is a split-half oracle: half the crowd predicting the other half. The gap between them is +0.2028 AP [+0.1698, +0.2342, domain-clustered], and three findings structure it. First, the gap is semantic: position and length features recover 5% of it. Second, frontier language models reach 35-53% of it zero-shot -- far above classical baselines, far below the crowd; a state-of-the-art prompt compressor (LLMLingua-2) lands below the floor, indistinguishable from random selection. Third, an unweighted cross-vendor fusion of five frontier r

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

What Could the Agent See at 19:05? Generating Temporal Enterprise Scenarios from Real Research and Replaying Them to Evaluate Agents

arXiv:2608.01042v1 Announce Type: cross Abstract: Enterprise AI agents act across many apps whose data changes continuously, so an answer is correct only relative to what data existed and who could see it at the moment it was asked. Offline evaluation today grades against a single static snapshot, effectively the end of the episode. So, it can only evaluate one situation, the final one, even though every earlier moment of the episode is a different situation that invites its own realistic questions with its own correct answers. Recreating each of those moments as a separate snapshot would mean re-provisioning a whole tenant per instant, which is prohibitively costly; and even a single snapshot leaks future state hidden inside records and cannot represent the multi-app, time-ordered way real work happens. Our system closes two gaps at once: it generates a realistic, persona-driven, temporally-evolving enterprise world from real research, and replays that world at any chosen moment to ev

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

MIDAL: Math Image Descriptions for Accessible Learning

arXiv:2608.00868v1 Announce Type: cross Abstract: Many open educational resources are lacking in accessibility, especially in-depth image descriptions. In subjects like Science and Mathematics, however, it can be particularly difficult to write image descriptions since there can be many complicated expressions and names depending upon the course level. To help fill that gap in a small way, we introduce Math Image Descriptions for Accessible Learning (MIDAL), a math image-description dataset of 2,020 mathematical images spanning multiple educational levels, to aid in training vision language models to create image descriptions following accessibility best practices. We hope MIDAL is a valuable resource in enhancing the conversation and innovation regarding accessibility of STEM content in higher education. This dataset is however not just limited in math description generation but can also be used to fine-tune language models that can have improved mathematical reasoning and answers.

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

XR-PRISM: Data-Driven Privacy and Risk Impact Scoring Metric for Extended Reality in Healthcare

arXiv:2608.00826v1 Announce Type: cross Abstract: Extended Reality (XR) technologies are transforming healthcare through immersive training, remote consultation, and patient rehabilitation. However, their extensive sensing capabilities and complex data pipelines introduce distinct security, privacy, and safety risks. Existing research lacks a unified quantitative framework for assessing and prioritizing these risks. We review 65 peer-reviewed studies on XR security and privacy published from 2017 to 2024, synthesizing a four-layer threat taxonomy consisting of Device, Network, User, and Cloud layers, along with a corresponding catalog of defenses. Building on this analysis, we introduce XR-PRISM, a six-factor weighted Privacy and Risk Impact Scoring Metric that integrates threat likelihood, system vulnerability, attack surface, safety impact, privacy impact, and control effectiveness into a single actionable risk score. Our analysis shows that more than 70% of the identified countermea

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

SoniSpeech: A Large-Scale Open-Vocabulary Tri-Modal Dataset for Wearable Silent Speech Interfaces

arXiv:2608.00803v1 Announce Type: cross Abstract: Wearable silent speech interfaces (SSIs) are limited to small, closed vocabularies. Approaches achieving larger vocabularies require obtrusive hardware such as facial electrodes. We present SoniSpeech, the first large-scale, open-vocabulary, trimodal dataset for wearable SSI using acoustic-sensing eyewear. It contains 34 hours across 18,000 utterances with three synchronized modalities: ultrasound echo profiles, voiced audio, and frontal video, in both voiced and silent modes. The corpus draws from the SODA dialogue dataset, providing contemporary conversational English with 5,356 unique words and full phoneme coverage. A CTC-based ResNet-34 baseline achieves 26.3% word error rate (WER) on open-vocabulary silent speech recognition, the first benchmark for this task. Dataset is available at https://doi.org/10.7298/xjjr-9m85

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

ORCESTRA: VLM-driven Visual Robot programming in Mixed Reality

arXiv:2608.00775v1 Announce Type: cross Abstract: ORCESTRA is a mixed-reality system for programming robot digital twins through no-code waypoint teaching and language-guided control. In a passthrough mixed-reality workspace, users place robot twins on real surfaces, teach trajectories, save robot-relative episodes, or issue spoken/typed commands that a vision-language model converts into structured digital-twin plans. Both interaction modes share a backend for metric grounding, embodiment-aware validation, preview, confirmation, and digital-twin execution. The system supports heterogeneous robot embodiments, including fixed-base manipulators, a mobile base, and a humanoid robot, demonstrating MR validation as a safety layer for language-guided robot programming before physical deployment.

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

From Chasing Ghosts to Missed Attacks: Perspectives and Perceptions of SOC Practitioners on LLM Integration, Risks, and Readiness

arXiv:2608.00672v1 Announce Type: cross Abstract: Security Operations Centers (SOCs) process large volumes of security events, requiring analysts to accurately detect and assess ongoing cyberattacks under time pressure. Recent advances in Large Language Models (LLMs) suggest potential benefits for security operations, yet their practical suitability for real-world SOC workflows remains poorly understood. To address this gap, we conducted 25 semi-structured interviews with SOC practitioners who had prior experience with LLMs, complemented by interactive scenarios to anticipate challenges and identify opportunities for the responsible integration of LLM-based tools into SOC workflows. We identified 15 LLM use cases grouped into six functional categories. While LLMs are valued for automating repetitive, low-level tasks such as report automation, practitioners rate high-impact tasks such as incident analysis as not yet feasible, reporting limitations in technical depth, context awareness,

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

A Context-Aware Cultural Heritage Guide Powered by LLMs

arXiv:2608.00549v1 Announce Type: cross Abstract: We present an extension of Triangolazioni (a Cultural Heritage webapp) to enrich curated content with context-dependent, external information provided by Large Language Models (LLMs) within a loosely-coupled architecture agnostic to the LLM. The system supports context-dependent information search and presentation within an architecture agnostic to the exploited LLM.

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Exploring Usability and Legal Practice: Insights from German Judicial Users of Digital Forensics

arXiv:2608.00541v1 Announce Type: cross Abstract: Digital forensics has become an integral part of modern criminal proceedings, yet its effective integration remains challenging because of increasing data volumes, evolving technologies, and complex interactions between technical and legal stakeholders. Although prior work has focused primarily on digital forensic tools and methods, its broader procedural and organizational context has received limited attention. Building on emerging perspectives inspired by usability research and human-centered security, we conceptualize digital forensics as part of a socio-technical system within criminal proceedings. We consequently investigate this perspective through a survey of 101 practitioners from the judiciary of the German federal state of North Rhine-Westphalia, including public prosecutors, judges, and digital forensic experts. The results indicate a strong demand for improved integration of digital forensics into workflows, enhanced cross-

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Ekova: A Personality-Support Agent for Self-Discovery Dialogue

arXiv:2608.00478v1 Announce Type: cross Abstract: Emotional Support (ES) systems have long optimized a single objective: alleviating the user's emotional distress in the moment. We argue that a complementary need, helping users see themselves more clearly, defines a distinct paradigm we call Personality Support (PS). PS is not counseling or clinical intervention: it targets cognitive clarity and self-articulation, not symptom relief or diagnosis. We instantiate this paradigm in three layers. First, we present DSD, a Chinese self-discovery PS Dataset of 8,590 samples collected through real longitudinal interaction across five minimal units, Coach, Warm, Tsukkomi, Real, and Gonzo. Second, we build DeepSupport, a multi-persona PS system trained with OrthoTune, a PS-tailored framework with style-specific adapters and a style-consistency regularizer. Third, we unify the five DeepSupport personas into Ekova, a persistent personality-support agent with a unified cross-session memory layer, su

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

On Defining Chart Types Boundaries

arXiv:2608.02512v1 Announce Type: new Abstract: What makes a Gantt chart? This question proved unexpectedly difficult to answer when we set out to build a design space for Gantt charts. Existing definitions, each shaped by their respective research goals, made different scope choices that we could not directly reconcile. We reasoned about what should and should not count as a Gantt chart, developing concepts and tools along the way. We distinguish features that are essential to a chart type's identity from those that can vary, and use these distinctions to map how chart types relate through what they share and lack. Applying these ideas to Gantt charts, radar charts, and table cartograms, we produce key insights on what boundary work reveals: definitions diverge for functional reasons, drawing boundaries exposes hidden structure in descriptive vocabulary such as feature entanglements, and scope choices shape how far findings can generalize. We came to understand that there is not a def

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Embodied Empathy: A Multimodal AR and LLM-Powered System for Self-Attachment Psychotherapy with Self-Initiated Humour

arXiv:2608.02283v1 Announce Type: new Abstract: The growing global demand for mental health support increasingly exceeds the supply of qualified practitioners, creating an urgent need for scalable digital interventions that can deliver meaningful emotional connection. In response, we present a novel multimodal application that operationalises the Self-Initiated Humour Protocol (SIHP) within a Self-Attachment Technique (SAT) framework. Our mobile application integrates customisable 3D childhood avatars, augmented reality, and an LLM-driven virtual therapist capable of automated emotion mirroring. An eight-day user study (N=16) indicates the system's feasibility and improvements in self-reported mood. Results show that personalised avatars and text-to-speech output strengthen emotional bonding and perceived empathy. Although emotion mirroring boosts engagement, its effectiveness depends heavily on classification accuracy and animation intensity. Moreover, findings indicate a shift in use

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

ManyFold: A Design Exploration of Data Visualization on Foldable Mobile Devices

arXiv:2608.02232v1 Announce Type: new Abstract: With this work, we explore the unique potential of data visualization on novel foldable mobile devices (foldables). Even though foldables are already commercially available, there is limited knowledge of how to leverage their distinct characteristics for visualization. This gap will only grow as their form factors become increasingly diverse. To address this, we use a two-step approach. First, we present a device-centered design space, structured around physical and usage properties of foldable devices. Second, we introduce a conceptual framework that investigates visualization on foldables from four complementary perspectives: More Displays - distributing multiple views to leverage additional display space; More Shapes - mapping visualizations to spatial fold states; More Interactions - coupling visualization tasks and folding interactions; and More States - enabling responsive visualization through folding. We complement the design spac

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

From Information to Delegation: Mapping Human-AI Financial Decision Making

arXiv:2608.02100v1 Announce Type: new Abstract: As AI increasingly participates in human decision making, understanding how decision-making authority is distributed between humans and AI has become a fundamental behavioural question. We introduce a behavioural measurement framework combining intent and delegated decision authority to quantify what consumers seek from AI and how much decision-making authority they assign to it. Applied to 1.5 million real-world ChatGPT and Gemini interactions from 6,304 users in the United States and India, we find that financial services are already a substantial AI use case. Consumers overwhelmingly use AI to retrieve information and shape financial judgement, while delegation of financial execution remains rare. By shifting attention from conversation topics to delegated decision authority, this work establishes a behavioural baseline for measuring the transition to increasingly agentic AI.

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

HaptoFlow: High-Fidelity Real-Time Vibrotactile Generation via Flow Matching for Virtual Reality

arXiv:2608.01974v1 Announce Type: new Abstract: Haptic feedback is widely employed to enhance immersion in Virtual Reality (VR) environments. However, designing haptic stimuli that cover diverse interaction conditions remains a significant scalability challenge. Data-driven haptic generation has emerged as a promising approach, yet existing models face an inherent trade-off between waveform expressiveness and inference responsiveness, which becomes increasingly critical as training data grow in scale and diversity. To address this challenge, we propose HaptoFlow, a vibrotactile generative model based on Flow Matching, designed for interactive real-time haptic rendering in VR. Flow Matching learns a continuous vector field that transforms a base distribution into the target data distribution, enabling efficient representation of complex haptic data distributions and thereby facilitating both high-quality generation and computational efficiency. We train HaptoFlow conditioned on material

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

3D Gaussian Splatting and Mesh-Based Digital Twins: An Exploratory Study for Virtual Reality Tourism

arXiv:2608.01969v1 Announce Type: new Abstract: Digital Twins (DTs) are increasingly used for immersive experiences in virtual tourism. Virtual Reality (VR) enables remote visits to replicated locations for promotional purposes or access to fragile and rural cultural heritage sites. However, developing high-fidelity DTs of tourist destinations is costly, due to the manual creation of 3D environments. Novel 3D rendering techniques, such as 3D Gaussian splatting (3DGS), pose a promising approach to creating immersive experiences. This study investigates the user experience (UX) of a 3D-mesh-based scene and a 3DGS-based scene within a VR tourism application. In a laboratory study, 20 participants engaged with both versions and rated UX, cybersickness, presence and affect through standardized questionnaires. A custom questionnaire was created to measure the perception of the DTs. The collected data suggests that both versions were enjoyed and induced positive affect, with the Mesh version

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Emotional Expression in Persuasion by Quadruped Virtual Agents: Toward Cross-Species Design Patterns

arXiv:2608.01895v1 Announce Type: new Abstract: Persuasive technologies increasingly use virtual agents to influence attitudes and behavior, but research has focused mainly on humanoid agents. The persuasive design of non-humanoid, quadruped agents remains underexplored, and it is unclear whether emotional expression works consistently across animal species or whether species-specific motion is necessary. We developed virtual dog, cat, and horse agents and compared three behavioral conditions: species-specific behavior, shared behavior across species, and a bark-only baseline. Participants completed everyday tasks involving trash disposal, feeding, and refraining from smartphone use. We evaluated intention understanding, behavioral intention, actual behavior, psychological reactance, discomfort, familiarity, and agent acceptance. In several task contexts, the bark-only baseline produced lower intention-understanding and behavioral scores than the expressive conditions. Emotional expres

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Reassessing the Feasibility of PPG-Based Non-Invasive Blood Glucose Level Estimation

arXiv:2608.01820v1 Announce Type: new Abstract: Non-invasive blood glucose level (BGL) estimation from photoplethysmography (PPG) holds great promise for wearable health monitoring, but results across studies are hard to compare due to inconsistent datasets, data leakage, and non-standardized evaluation metrics. We present the first reproducible, extensible evaluation pipeline and use it to reassess five representative PPG-based BGL methods on published datasets under three increasingly strict data-split protocols: random window-level, participant-aware, and leave-some-participants-out (LSPO). Models appeared competitive under random splitting but collapsed under participant-aware and LSPO evaluation, with nearly all yielding near-zero or negative R$^2$ values comparable to a mean-prediction baseline. Critically, across every model and split, over 90% of predictions fell within clinically acceptable zones (Clarke Error Grid A+B), including the baseline. This reveals a fundamental disco

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Beyond Score-Based Gamification: Designing Spatiotemporal and Musical Experiences for VR Neck Rehabilitation

arXiv:2608.01688v1 Announce Type: new Abstract: Pain-related anxiety and fear of movement are major barriers to adherence and therapeutic outcomes in rehabilitation exercises for chronic neck pain. Virtual reality enables the design of immersive experiences that can transform repetitive therapeutic movements into engaging and emotionally supportive interactions. In this exploratory work, we investigate how experience-oriented gamification can reduce anxiety and improve user experience during VR-based neck range-of-motion exercises. We introduce two novel interaction paradigms that embed therapeutic neck movements within multisensory VR experiences. The first paradigm, Spatiotemporal Progression, couples head-tracked trajectories with environmental progression in a tropical island setting, where movement segments dynamically transform time of day, weather, and spatial location as experiential rewards. The second paradigm, Musical Interaction, maps movement segments to meditative music n

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

CellPrism: A Visual Analytics System for Exploring AI-Driven Virtual Cells in Drug Discovery

arXiv:2608.01669v1 Announce Type: new Abstract: Gene perturbation analysis plays a critical role in drug discovery by enabling researchers to investigate how interventions on specific genes influence global gene expression patterns within cells. Recent advances in artificial intelligence-driven virtual cell models have made it possible to predict gene expression outcomes for a wide range of perturbation strategies in silico, substantially reducing reliance on costly and time-consuming biological experiments. However, effectively exploring and interpreting the high-dimensional perturbation spaces produced by these models remains challenging because of the combinatorial nature of perturbations and the complex cell-specific gene expression responses they generate. In this work, we present CellPrism, a visual analytics system designed to support the systematic exploration of gene perturbation strategies for drug discovery. Specifically, CellPrism integrates clustering-based overviews to su

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

You Cannot Optimize What You Cannot Measure: Multitasking Evaluation as the Missing Foundation of AI-Mediated Heads-Up Interaction

arXiv:2608.01656v1 Announce Type: new Abstract: AI-mediated heads-up augmented reality (AR) replaces fixed interfaces with dynamically adapting ones that decide what information to present, in what form, and when, based on a continually changing context that cannot be fully anticipated beforehand. Although it remains an interface, its behavior over time is only partially specified at design time. We argue that this shift requires a corresponding change in evaluation: from snapshots to trajectories. A fixed interface is evaluated in a snapshot --- one context, one session, one set of task-performance metrics. A fluid interface must be evaluated over a trajectory --- a sequence of contexts with transitions, sampled from the distribution the interface will actually encounter, and tracked long enough for user trust to form, evolve, and potentially deteriorate. Drawing on the literature for heads-up AR multitasking enabled by optical see-through head-mounted displays (OST-HMDs), we find tha

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

FedWorld: Scope-Aware Federation of Agent World Models

arXiv:2608.01561v1 Announce Type: new Abstract: Large language model (LLM) agents learn world dynamics from local interaction experience to support subsequent planning and action selection. However, the experience available to a single client is often incomplete, which motivates sharing knowledge across clients. Existing federated methods mainly aggregate model parameters, while agent memory-sharing methods commonly pool trajectories, memories, or rules without checking whether they remain valid for each client. This assumption is problematic because the same abstract action may produce different effects under different policies, environments, or exception conditions. Consequently, a rule supported by most clients may overwrite correct knowledge held by a minority client. To address this problem, we propose FEDWORLD, a scope-aware federated world-model protocol that exchanges structured abstract transition rules. Each client converts private transitions into normalized rules, and the s

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

A Data-Centric Perspective on Tree Visualizations

arXiv:2608.01477v1 Announce Type: new Abstract: Tree visualization (TreeVis) techniques span diverse designs. Existing taxonomies organize them by visual characteristics such as layout dimensionality, edge representation, and node alignment. However, this visual-centric perspective can obscure structural similarities and make it difficult to determine whether differences arise from data structures or visual encodings. We investigate TreeVis techniques from a data-centric perspective grounded in Prepared Tables, the final data state prior to visual encoding. Using TreeVis.net, we curate 133 two-dimensional techniques and characterize each by the object records and attribute roles required before encoding. Our analysis shows that the corpus is more concentrated at the prepared-data level than a visual reading would suggest. The techniques collapse to a small set of recurring object combinations and schemas. Many techniques across TreeVis representation categories share the same schema, s

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

PartInteractor: Intent-Driven Part-Aware 3D Authoring for Continuous Co-Creation in XR

arXiv:2608.01335v1 Announce Type: new Abstract: As Extended Reality (XR) evolves into an immersive computing medium, interactive 3D authoring becomes essential for creative and functional workflows. However, existing generative XR systems produce monolithic outputs lacking explicit semantic structure, limiting post-generation control. We introduce PartInteractor, a representation-to-interaction framework that investigates how semantic part hierarchies can be incorporated into generative XR authoring, and exposed as first-class, directly manipulable units, turning one-shot prompt-to-object generation into continuous component-level co-creation. PartInteractor supports speech, sketch, and image inputs, integrating an LLM interpreter with a retrieval-generation strategy to scaffold user intent prior to 3D generation. Instead of producing monolithic objects, our system generates semantically decomposed 3D assets with explicit part hierarchies, enabling rich component-level interaction over

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Collascope: Supporting Serendipitous Asset Exploration for Collage-Based Storytelling

arXiv:2608.01267v1 Announce Type: new Abstract: Collage-based storytelling requires visual elements that support emerging narratives and inspire creative reinterpretation. Existing tools, however, rely largely on keyword- and image-based retrieval, offering limited support for serendipitous exploration beyond existing assets. We introduce Collascope, an interactive system that helps creators (1) concretize story intent with interactive element groups, (2) expand the exploration space based on concepts or cutouts towards conceptual and visual dimensions, and (3) develop grounded, traceable ideas in parallel with collage composition. Collascope's attribute-aware visual retrieval method, instantiated with collage-relevant visual dimensions, enables creators to retrieve cutouts through dimension-specific visual projections rather than holistic similarity. In a within-subject study (N=12) against a conventional search baseline, our participants used unexpected results and even gaps in the a

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Rethinking PPG-based Sleep Staging: Datasets, Metrics, and Benchmarks

arXiv:2608.00943v1 Announce Type: new Abstract: Automated sleep staging assigns discrete stage labels to successive time epochs throughout an overnight recording; conventionally each window spans at least 30 seconds, reflecting the minimum temporal resolution of the clinical scoring standard. Wearable photoplethysmography (PPG) has attracted sustained interest as an ambulatory alternative to laboratory-based polysomnography, which relies on electroencephalography (EEG) and other recording modalities that are impractical outside clinical environments. Yet PPG-based staging trails EEG-based methods by a substantial margin, and we argue this gap largely reflects a mismatch between signal and task. Within a stable stage, PPG's inter-stage feature differences are more subtle than those in EEG; yet at stage boundaries, PPG's principal cardiovascular features, heart rate variability and pulse morphology, shift sharply within seconds. The conventional practice of assigning one label to each 30

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

The Assistant Erased You: Measuring Loss of Authorship Signals in AI-Mediated Communication

arXiv:2608.00926v1 Announce Type: new Abstract: Research on AI-mediated communication has examined how AI assistance shapes interpersonal perceptions and reduces stylistic diversity across users. We ask a complementary question at the individual level: after a message is rewritten by an AI writing assistant, can its author still be distinguished from others? We introduce the Idiolect Erasure Rate (IER), defined as the reduction in authorship-attribution accuracy following AI-assisted rewriting. We evaluate IER on three pre-generative-AI corpora using a stylometric model and the authorship-specific LUAR model. Heavy rewriting substantially weakens authorship signals in personal blogs and workplace email, reducing LUAR attribution by as much as 66.5 percentage points, but has a much smaller effect on topic-structured news, where topic remains predictive of authorship. Additional analyses suggest that rewriting produces stylistic convergence despite substantial semantic overlap, and that

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Who's That Player?: Externalizing Query Interpretation in Spoken XR Sports Interaction

arXiv:2608.00876v1 Announce Type: new Abstract: XR sports viewing enables spectators to follow play from immersive, spatially anchored perspectives while accessing contextual analytics directly within the scene. In such settings, speech offers a practical interaction modality because text entry and menu navigation can interrupt attention during fast-paced gameplay. However, spoken queries are often underspecified: viewers may omit which player, time period, field location, or metric they intend. When systems resolve these ambiguities implicitly, their assumptions remain hidden, making misinterpretations difficult to notice and correct (repair). We investigate how externalizing a system's interpretation of spoken queries can support inspection and correction of such misunderstandings in XR sports viewing. Through a formative study, we identified four recurring ambiguity types (referential, spatial, temporal, and metric) that characterize ambiguous spoken queries in this context. We deve

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Continuous Face Authentication on Mobile and Desktop Platforms: A Comparative Study

arXiv:2608.00763v1 Announce Type: new Abstract: Personal devices hold sensitive data and provide access to sensitive services. Conventional personal device authentication verifies users' identity only at the moment access is granted. An unlocked device may be accessed by an unauthorized person if the user stops using the device without locking it, or if another person takes over. Continuous authentication addresses this gap. This paper investigates how device type and usage conditions influence continuous mobile face authentication with an InsightFace-based approach with temporal trust decay. We evaluate the approach with mobile and desktop recordings with different head directions and lighting conditions. We also evaluate recordings from everyday mobile device use without predefined tasks. The results show that device type alone has little impact, while different usage conditions do have impact on the authentication performance. Results also show that everyday mobile device use is in

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Me and My Bot: What Users Talk About in AI Companion Communities on Reddit

arXiv:2608.00748v1 Announce Type: new Abstract: AI companion communities on platforms such as Reddit are widely characterized as spaces where users discuss their relationships with AI bots. This study examines whether and how that characterization holds, guided by the Synthetic Resonance framework's claim that human-AI relationships can carry genuine relational meaning for the user. Multiple LLMs were employed to code 5,504 Reddit posts from eight AI companion communities for relationship focus, primary topic, and users' emotional valence. Although search terms were weighted toward relational and attachment language, only 45% of posts concerned the user's own relationship with their bot. Posts about users' own bots differed markedly from posts about bots in general in both topic and emotional expression, with 85% of general-bot posts containing no user emotion language compared to 33% of own-bot posts. Among the 970 posts that were relationally focused, companionship and romance each a

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

bFaaaP: An Inclusive, Head-Angle Piano-Pedal Interaction that Quantitatively Reproduces a Pianist's Intended Pedalling -- Foot-Free, for Acoustic and Electronic Pianos

arXiv:2608.00633v1 Announce Type: new Abstract: Expressive piano performance depends on the sustain (damper) pedal, operated by foot, excluding players who cannot readily use their feet: wheelchair users and others with lower-limb impairments, small children, and some elderly or disabled players. We present bFaaaP (barrier-Free assist as a Pedal), an inclusive, foot-free interaction that operates the pedal from the angle of the player's head: a smartphone tracks head pose with on-device augmented-reality (AR) face tracking and streams a compact command over Bluetooth Low Energy (BLE) to a pedal device. Supported by patent examination, our central claim is not the head-to-pedal architecture (anticipated by prior art) but a quantitative, user-tunable control law -- the patentable "key" to a natural, expressive result: the player presets a small angular dead-zone (offset 3-10 degrees) and a multiplier (10-50), which together fix a secondary, pre-adjustable response speed that reproduces t

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

SkinSpline: A Body-Attached Skeleton-Supported Haptic Interface for Continuous Skin Deformation through Physical Interpolation

arXiv:2608.00496v1 Announce Type: new Abstract: We present SkinSpline, a body-attached skeleton-supported haptic interface that renders continuous skin deformation through physical interpolation of sparse mechanical actuation. SkinSpline combines a low-resolution array of rack-and-pinion linear actuators with an elastic interlocking skeleton that transforms discrete actuator motions into smooth surface deformation, enabling continuous cutaneous feedback without dense actuator arrays. The system includes a modular hardware architecture, a configurable control pipeline, and a visual interface supporting real-time configuration and actuation. We demonstrate SkinSpline through multiple scenarios, including wave rendering, video-synchronized rhythmic touch, visually driven water-wave feedback in VR, and sensor-based remote touch reproduction. SkinSpline explores an alternative approach to continuous on-body haptic rendering by leveraging structural coupling between sparse actuation and defo

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Revibing Code from Papers: Reimplementing HCI Artifacts

arXiv:2608.00450v1 Announce Type: new Abstract: Software artifacts for most technical HCI research projects are unavailable. The lack of access to these imposes limits on academic knowledge production. It is difficult to: extend or reuse research artifacts; use strong baselines in evaluating follow-up work; and perform replication or reproducibility research. In this work, we demonstrate the potential of new agentic AI technologies to revibe interactive software: reimplement systems directly from research papers. To measure the success of the approach, we describe a revibeability metric. By revibing recent research papers from UIST, and interviewing their original authors, we demonstrate the plausibility (and limitations) of revibed system. The results are encouraging. In many cases producing code suitable for strong baseline use. We argue that this may represent a fundamental shift in how we produce, use, and evaluate research artifacts in the technical HCI community.

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Seeing Through the Forecast Clutter: Communicating Climate Forecast Distributions with Weighted Multiple Forecast Visualizations

arXiv:2608.00433v1 Announce Type: new Abstract: Forecasts often diverge because different models make varying assumptions to account for underlying uncertainty. Readers who consume forecasts may wish to survey the shape and spread of these multiple forecasts to get a full account of the different predictions. One approach to visualizing multiple forecasts is through Confidence Interval (CI) plots. However, while the summative CI plots can communicate uncertainty of an ensemble, they obscure attributes of individual forecasts that can lead to inaccurate perceptions of the distribution of these forecasts (e.g., implying a normal distribution when non-existent). To address this challenge, we investigate the use of multiple forecast visualization (MFV) in communicating nuanced forecast distributions through two preregistered experiments using climate forecast data. In Experiment 1 (480 participants), we compared how well MFV and CI plots can represent the distribution of multiple forecasts

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Visualizing Placement Proposals for Window Arrangement in Mixed Reality: A Comparative User Study

arXiv:2608.00403v1 Announce Type: new Abstract: Adaptive mixed reality (MR) interfaces typically optimize window layouts on behalf of the user, with limited consideration for individual preferences. A promising alternative keeps users in the loop by presenting layout proposals for them to select from, but how these proposals should be visualized remains underexplored. We compare three proposal-visualization techniques for window placement, Situated Icon Preview, Situated Window Preview, and 3D Preview, against a Manual Positioning baseline. The techniques differ in level of detail and degree of interaction-space context. In a within-subjects user study, 24 participants completed a multi-stage trip-planning task in VR, individually placing seven sequentially introduced windows using each technique. We thereby focus on single-window placement under predefined proposal positions. We measured layouting time, number of layout changes, task load, user experience, and preference, complemented

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

MolecularCanvas: LLM-assisted Small-Molecule Drug Discovery via Structure-Guided Constraints

arXiv:2608.00393v1 Announce Type: new Abstract: Small-molecule drug discovery relies on iterative molecular optimization, where chemists repeatedly modify candidate compounds to balance multiple competing properties such as efficacy, toxicity, and solubility. Recent advances in generative AI (GenAI) have shown promise in accelerating this process by automatically proposing new molecular structures or targeted modifications. However, existing GenAI-based molecular design tools remain poorly aligned with experts' real-world workflows. Specifically, they offer limited support for specifying structure-level modification intents on molecules, provide insufficient transparency into model-generated modifications, and lack integrated support for downstream property evaluation with external computational tools. To address these challenges, we introduce MolecularCanvas, an interactive system that enables users to iteratively construct an optimization context by integrating high-level goals, stru

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Dynamic Surveys: Using LLMs to Blend Qualitative Depth,Quantitative Structure, and Collaborative Interaction

arXiv:2608.00357v1 Announce Type: new Abstract: Surveys are a powerful tool for collecting data and eliciting insights on social phenomena, and are critical in product design, marketing, scientific research. However, traditional open-ended and closed-ended question formats limit researchers' ability to capture data that combines both the richness of qualitative insights and the analytical rigor of quantitative data. To address these problems, we propose Dynamic Surveys, a survey platform that uses Large Language Models (LLMs) to dynamically cluster qualitative responses in real time and to elicit quantitative ratings and rankings on those clusters and qualitative reflections on how their views compare to broader respondent trends, especially helpful in early-stage or exploratory research settings. This process generates a report showing survey creators and respondents the clustered responses as well as each cluster's rank, rating distribution, and follow-up reflections. To evaluate Dyn

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Read, Critique, or Sketch? Investigating Alternative Visualization Literacy Assessment Modalities

arXiv:2608.00330v1 Announce Type: new Abstract: Visualization literacy is a multifaceted construct encompassing skills and competencies, such as decoding data, constructing charts, and identifying design flaws. Yet, assessments of these competencies has been primarily constrained to multiple choice assessments that target lower-order skills, such as chart comprehension. As a result, they often exhibit ceiling effects (i.e., even modestly skilled individuals commonly score near the top of the scale), and do not provide enough information about an individual's higher-order skills (e.g., applying external knowledge, formulating critiques, and designing visualizations). To close these gaps, we develop and investigate two web-based qualitative assessments for testing the critique and design aspects of visualization literacy through online think-aloud critique and sketching of visualization designs based on data and a prompt. We compare performance on our assessments to two established visua

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Theory, Experience, and Instinct: A Glimpse Into AAA Game Processes and How UX Leaders Navigate Pre-Production

arXiv:2608.00313v1 Announce Type: new Abstract: Foundational decisions shape a project's long-term trajectory, a dynamic that becomes especially evident in the inherent complexity of game pre-production. However, academic frameworks often see limited uptake at this stage, as they do not readily map onto industry contexts, production constraints, and cross-functional workflows. To better understand how design decisions are made in practice, we conducted interviews with 15 UX leaders from the AAA (triple-A) games industry. Our findings show that early UX decisions emerge from a dynamic blend of theory, experience, and intuition. In cross-functional structures (such as strike and competency teams), UX leaders collaboratively align player needs, technical feasibility, and creative vision. These decision-making processes involve translating academic concepts into production-ready insights, codifying experiential knowledge into reusable practices, and relying on informed intuition amid uncer

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

ReVoicer: Conversational Voice Annotation for Human-Centered, LLM-Assisted Peer Review

arXiv:2608.00299v1 Announce Type: new Abstract: We present ReVoicer, a prototype system that supports peer reviewers by letting them converse with a paper as they read it. The reviewer highlights a passage and speaks (or types) a train-of-thought comment. A large language model then cleans the comment using the surrounding prose as context, tags it by comment type, and anchors it to the passage. After the reviewer finishes reading, ReVoicer checks the accumulated notes against a venue-specific rubric, reports coverage gaps, and drafts a review composed only from the reviewer's own comments, written to a style guide distilled from the reviewer's past reviews. The system introduces no critiques of its own. We describe the system's design rationale and implementation, and we outline plans for future evaluations. With the ISMAR community, we will gather feedback and discuss the system design and ideas for additional features and evaluations.

Source ↗
technology Tue, 04 Aug 2026 00:00:00 -0400
arXiv cs.HC

Textro: A Prototyping Toolkit for Solderless and Chipless Smart Textile Interfaces

arXiv:2608.00294v1 Announce Type: new Abstract: In this paper, we present Textro, a prototyping toolkit for designing, fabricating, and testing solderless and chipless smart textile interfaces. Unlike prior approaches that rely on rigid components or soldered connections, Textro enables users to build functional textile interfaces using only readily available materials and tools. The toolkit integrates three parts: (1) a web-based design environment for importing sewing patterns, defining sensing elements, and automatically generating optimized component and circuit designs based on empirical experiments; (2) a fabrication pipeline that generates fabrication files for embroidery and cutting machines, with embroidery optimized for one-stroke continuous stitching paths and components assembled through glue-based attachment methods via capacitive coupling; and (3) a reader device and software for wirelessly retrieving sensor data and visualizing real-time sensor signals. We demonstrate Te

Source ↗
technology Tue, 01 Sep 2026 00:00:00 -0400
arXiv cs.HC

How Coding Agents Fail Their Users: A Large-Scale Analysis of Developer-Agent Misalignment in 20,574 Real-World Sessions

arXiv:2605.29442v2 Announce Type: replace-cross Abstract: AI coding agents increasingly act directly within software environments, yet existing analyses of their failures rely on benchmark trajectories that miss how developers actually experience misalignment. We present an observational study of 20,574 coding-agent sessions from 1,639 repositories across IDE and CLI workflows. We operationalize misalignment as a breakdown made visible through developer pushback, and annotate each episode along four axes: form, cause, cost, and resolution. We identify seven recurring forms, spanning how agents read projects, interpret developer intent, follow rules, bound their actions, implement and execute code, and report progress. 90.50% of episodes impose effort and trust costs rather than irreversible system damage, yet 91.49% of visible resolutions still require explicit user correction. Misalignment patterns also differ across IDE and CLI settings, persist across adjacent sessions, and shift ov

Source ↗
technology Tue, 01 Sep 2026 00:00:00 -0400
arXiv cs.HC

Active Inference with People: a general approach to real-time adaptive experiments

arXiv:2603.29003v2 Announce Type: replace-cross Abstract: Adaptive experiments optimize their design throughout data collection, which can bring substantial benefits compared to conventional experimental settings. Potential applications include, among others, computerized adaptive testing (when selecting informative tasks in ability measurements), adaptive treatment assignment (when searching for experimental conditions maximizing certain outcomes), and active learning (when choosing optimal training data for machine learning algorithms). However, implementing these techniques in real time poses substantial computational and technical challenges. In this paper, we introduce a practical and unified approach to real-time adaptive experiments that can encompass these scenarios across textual, visual, and audio tasks. Our strategy combines active inference, a Bayesian framework inspired by cognitive neuroscience, with Pyro, a probabilistic programming library, and PsyNet, a modular Python

Source ↗
technology Tue, 01 Sep 2026 00:00:00 -0400
arXiv cs.HC

Pearmut: Human Evaluation of Translation Made Trivial

arXiv:2601.02933v4 Announce Type: replace-cross Abstract: Human evaluation is the gold standard for multilingual NLP, but is often skipped in practice and substituted with automatic metrics because it is notoriously complex and slow to set up with existing tools with substantial engineering and operational overhead. We introduce Pearmut, a lightweight yet feature-rich platform that makes end-to-end human evaluation as easy to run as automatic evaluation. Pearmut removes common entry barriers and provides support for evaluating multilingual tasks, with a particular focus on machine translation. The platform implements standard evaluation protocols, including DA, ESA, and MQM, and is extensible to support new protocols. It features document-level context, absolute and contrastive evaluation, attention checks, ESAAI pre-annotations and both static and dynamic assignment strategies. Pearmut enables reliable human evaluation to become a practical, routine component of model development and

Source ↗
technology Tue, 01 Sep 2026 00:00:00 -0400
arXiv cs.HC

TxSum: User-Centered Ethereum Transaction Understanding with Micro-Level Semantic Grounding

arXiv:2512.06933v4 Announce Type: replace-cross Abstract: Understanding the economic intent of Ethereum transactions is critical for user safety, yet current tools expose only raw on-chain data or surface-level intent, leading to widespread ``blind signing'' (approving transactions without understanding them). Through interviews with 16 Web3 users, we find that effective explanations should be structured, risk-aware, and grounded at the token-flow level. Motivated by these findings, we formulate TxSum, a new domain-grounded NLP task for DeFi transaction explanation, and construct a dataset of 187 complex Ethereum transactions with 2,375 token-flow annotations and transaction-level summaries. We further introduce MATEX, a grounded multi-agent framework for high-stakes transaction explanation. It selectively retrieves external knowledge under uncertainty and audits explanations against raw traces to improve token-flow-level factual consistency. MATEX achieves the strongest overall explan

Source ↗
technology Tue, 01 Sep 2026 00:00:00 -0400
arXiv cs.HC

When Chatbots Accommodate: Auditing the Response Policies of AI Companions in Vulnerable Conversations

arXiv:2606.04431v2 Announce Type: replace Abstract: Millions turn to AI companion chatbots during loneliness, grief, and personal crises. How these companion platforms respond in such moments can shape the trajectory of a user's vulnerable state. Yet existing model audits evaluate reactions to pre-defined crisis prompts and miss the response policy that governs sustained real-world interaction. We address these gaps with two key contributions. First, we introduce the AI Companion Vulnerability-Response Taxonomy, a grounded, paired taxonomy of user vulnerability and chatbot response designed for analyzing extended companion chatbot interactions. Second, we apply Maximum Causal Entropy Inverse Reinforcement Learning to ~47k turns of real-world user conversations with GPT-4.1, Character.AI, and Replika to infer each platform's short-horizon response policy: the probability of each response category given the user's current vulnerability state. Our findings reveal distinct response profile

Source ↗
technology Tue, 01 Sep 2026 00:00:00 -0400
arXiv cs.HC

AwareLLM: A Proactive Multimodal Ecosystem for Personalized Human-AI Collaboration to Enhance Productivity

arXiv:2605.09625v3 Announce Type: replace Abstract: Information workers' productivity is significantly influenced by their cognitive states and physiological responses. AI assistants such as ChatGPT, Copilot, and others have become integral components of knowledge-intensive workplaces. These AI assistants utilize pre-defined user preferences and chat interaction histories, thus confining themselves to reactive exchanges, lacking sufficient adaptability. Consequently, they fail to cater to individual user preferences and are unable to adapt to their psychophysiological states, diminishing potential productivity gains. To bridge this gap, we introduce AwareLLM, a novel multimodal framework that integrates egocentric vision, pupillometry, eye-gaze tracking, posture detection, heart activity, and the inferencing capabilities of large language models (LLMs) to create a proactive and context-aware ecosystem. AwareLLM dynamically adapts to users' psychophysiological states while analyzing tem

Source ↗
technology Tue, 01 Sep 2026 00:00:00 -0400
arXiv cs.HC

LatentGandr: Visual Exploration of Generative AI Latent Space via Local Embeddings

arXiv:2604.19953v2 Announce Type: replace Abstract: Generative AI has demonstrated significant potential in creative design, enabling the rapid generation of visual content and imaginative concepts. Although deep AI models achieve effective featurization in the latent space, navigating the space remains a challenge. Current techniques, such as GANSlider and SliderSpace, use multiple sliders to generate high-dimensional vectors in generative AI's latent space. Despite applying (global) PCA to reduce the number of sliders, these approaches struggle with scalability and usability as the number of control dimensions increases. In this paper, we introduce LatentGandr, a visual analytics technique that facilitates latent space exploration by extracting locally linear dimensions from embeddings in high-dimensional latent spaces. By analyzing the topology and local curvature of the embeddings, LatentGandr automatically identifies local neighborhoods and computes their principal components usin

Source ↗
technology Tue, 01 Sep 2026 00:00:00 -0400
arXiv cs.HC

The Double-Edged Sword of Open-Ended Interaction: How LLM-Driven NPCs Affect Players' Cognitive Load and Gaming Experience

arXiv:2604.10107v2 Announce Type: replace Abstract: This study examines how large language model-driven non-player characters (LLM-NPCs) affect players' cognitive load and gaming experience, with a particular focus on the underlying psychological mechanisms, differences across task scenarios, and the role of individual traits. Conducting a randomized between-subject experiment (N=130) in a self-developed game prototype "Campus Culture Week", we compared player interactions with LLM-NPCs and traditional pre-scripted NPCs across multiple interactive modules. The results showed that LLM-NPCs significantly increased players' cognitive load (p < .001), an effect mediated by factors such as expressive effort and response uncertainty. However, LLM-NPCs did not yield a statistically significant improvement in overall gaming experience (p = .195); while they positively influenced players' perceived autonomy, they exerted a negative influence on system usability and trust. The effects of LLM-NPC

Source ↗
technology Tue, 01 Sep 2026 00:00:00 -0400
arXiv cs.HC

Varifocal Displays Reduce the Impact of the Vergence-Accommodation Conflict on 3D Pointing Performance in Augmented Reality Systems

arXiv:2602.05129v2 Announce Type: replace Abstract: This paper investigates whether a custom varifocal display can improve 3D pointing performance in augmented reality (AR), where the vergence-accommodation conflict (VAC) is known to impair interaction. Varifocal displays have been hypothesized to alleviate the VAC by dynamically matching the focal distance to the user's gaze-defined target depth. Following prior work, we conducted a within-subject study with 24 participants performing an ISO 9241-411 pointing task under varifocal and fixed-focal viewing. Overall, varifocal viewing yielded significantly higher performance than the fixed-focal baseline across key interaction metrics, although the magnitude and even the direction of the benefit varied across individuals. In particular, participants' responses exhibited a baseline-dependent pattern, with smaller improvements (or occasional degradation) observed for those with better baseline performance. Our findings suggest that varifoca

Source ↗
Showing 751–800 of 1631 signals
← Prev Page 16 of 33 Next →