ABOUT THIS ISSUE

How was this newsletter synthesized?

Methodology

This newsletter is generated by an AI pipeline (leveraging Anthropic Sonnet 4.5 & Haiku 4.5) that processes the metadata and abstracts of every new arXiv HCI paper from the past week—87 this issue. Each paper is scored on three dimensions: Practice (applicability for practitioners), Research (scientific contribution), and Strategy (industry implications), with scores from 1-5. Papers passing threshold are grouped into topic clusters, and each cluster is summarized to capture what that body of research is exploring.

Selection Criteria

The pipeline builds a curated selection that balances high scores with topic diversity—and deliberately includes at least one 'contrarian' paper that challenges prevailing assumptions. This selection is then analyzed to identify key findings (patterns across multiple papers) and surprises (results that contradict conventional wisdom). A narrative synthesis ties the week's research together under a unifying frame.

Key Themes Discovered

Field Report: ai-interaction

Trust, Calibration, and Oversight

This cluster examines how humans evaluate, trust, and maintain control over AI systems in sustained interaction. Core tensions emerge: memory systems score high on benchmarks but fail to integrate naturally into conversation; harm detection requires multi-turn context that models struggle to process; and user-authored policies paradoxically reduce protection while preserving choice. Workers and older adults lack signals for when intervention is needed. The research prioritizes situated evaluation—measuring what matters in practice rather than aggregate performance—and explores how design friction, transparency, and worker agency can sustain meaningful human oversight as AI systems gain autonomy.

1/10