Hugging Face Daily PapersNan Li, Albert Gatt, Massimo Poesio1 min readpaperintermediate
Gaze as Evidence for Common Grounding: A Cross-Corpus Analysis of MapTask and MUNDEX
Summary
Cross‑corpus study of gaze behavior in two collaborative dialogue datasets (MapTask, MUNDEX) shows that task‑aligned references correlate with more task‑directed, less partner‑directed gaze, lower entropy and fewer transitions. Temporal gaze features (MapTask) and raw proportion features (MUNDX) modestly improve grounding prediction over baselines, but effects are small and diminish when aggregat…
- Mapped both corpora to a unified partner/task/away label set to compare gaze patterns.
- Aligned references (MapTask) and UND judgments (MUNDEX) exhibit more task‑focused gaze and reduced partner‑focused gaze.
- Gaze entropy drops and transition counts fall at moments of successful grounding.
- Temporal gaze features (e.g., fixation windows) give the best lift in MapTask; raw proportion features work best in MUNDEX.
Understanding how visual attention signals grounding can inform multimodal dialogue systems and assistive interfaces that need to infer shared understanding from limited cues.
6/10
