---
res:
  bibo_abstract:
  - |-
    <jats:title>Abstract</jats:title>
                      <jats:p>Investigating the real-time interplay between language and vision has traditionally involved a trade-off between experimental control and interactive setups. To address this challenge, we present a methodological framework that embeds the visual world paradigm (VWP) within a continuous, goal-directed human–robot interaction (HRI) task. By augmenting VWP in an HRI setup to produce standardized yet responsive verbal and non-verbal behaviours, the framework makes it possible to selectively examine comprehension, production, actions and their relation to visual attention in real-time. We demonstrate the utility of the framework through a proof-of-concept manipulating action structure, specifically path and manner, as well as other-initiated repair formats, namely open and restricted requests. The findings show that the setup reliably captures time-locked referential gaze dynamics following comprehension and preceding production processes. They further indicate that differences in action history systematically influence speakers’ pre-verbal cognitive processes and their subsequent linguistic choices. These findings suggest that HRI can provide a useful methodological middle ground to investigate language processes online in a tightly controlled setup without disrupting the interaction. More broadly, the framework offers a test bed for examining whether established psycholinguistic findings extend to HRI setups. At the same time, it provides HRI researchers with an empirical basis for developing adaptive agents informed by online language-vision interaction theories.</jats:p>@eng
  bibo_authorlist:
  - foaf_Person:
      foaf_givenName: Amit
      foaf_name: Singh, Amit
      foaf_surname: Singh
  - foaf_Person:
      foaf_givenName: Katharina J.
      foaf_name: Rohlfing, Katharina J.
      foaf_surname: Rohlfing
  bibo_doi: 10.1007/s10936-026-10296-z
  bibo_issue: '5'
  bibo_volume: 55
  dct_date: 2026^xs_gYear
  dct_isPartOf:
  - http://id.crossref.org/issn/0090-6905
  - http://id.crossref.org/issn/1573-6555
  dct_language: eng
  dct_publisher: Springer Science and Business Media LLC@
  dct_title: A Methodological Framework for Studying Online Language–Vision Interplay
    in Human–Robot Interaction@
...
