Does Local Video Understanding Transfer Across Encounters? The EgoGears Benchmark
아직 한국어판이 없어 영어 원문으로 표시.
초록 발췌
Embodied systems must make knowledge acquired during one encounter usable in another despite changes in viewpoint, motion, and illumination. Yet aggregate cross-video accuracy conflates failures of local perception with failures to preserve observation identity, establish correspondence, and compose evidence, obscuring whether local video understanding actually transfers.
초록에서 가져왔습니다. 요약을 준비 중입니다.