STARS: From Spatiotemporal Dynamics to Social Representations in Human-Robot Interaction
아직 한국어판이 없어 영어 원문으로 표시.
초록 발췌
Robot navigation in dynamic, human-centered environments requires socially-compliant decisions grounded in robust scene understanding. Recent Vision-Language Models (VLMs) exhibit promising capabilities such as object recognition, common-sense reasoning, and contextual understanding, capabilities that align with the nuanced requirements of social robot navigation.
초록에서 가져왔습니다. 요약을 준비 중입니다.