pith. sign in

arxiv: 2503.09511 · v1 · pith:HFWBYXM7new · submitted 2025-03-12 · 💻 cs.CL

TRACE: Real-Time Multimodal Common Ground Tracking in Situated Collaborative Dialogues

classification 💻 cs.CL
keywords tracemultimodalcollaborativecommongroundreal-timesituatedtracking
0
0 comments X
read the original abstract

We present TRACE, a novel system for live *common ground* tracking in situated collaborative tasks. With a focus on fast, real-time performance, TRACE tracks the speech, actions, gestures, and visual attention of participants, uses these multimodal inputs to determine the set of task-relevant propositions that have been raised as the dialogue progresses, and tracks the group's epistemic position and beliefs toward them as the task unfolds. Amid increased interest in AI systems that can mediate collaborations, TRACE represents an important step forward for agents that can engage with multiparty, multimodal discourse.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.