Video2STL: Grounding VLM-Generated Temporal Specifications for Robot Learning
日本語版は未提供のため、英語原文で表示しています。
要旨より
Video-based policy learning is particularly promising, as it illustrates target behaviors without requiring action annotations or embodiment-matched demonstrations. A central challenge is deciding what information should be transferred from the video to the robot.
要旨より。当社による要約は作成中です。