PhysWAM: Physically Consistent World Action Model for Autonomous Driving
아직 한국어판이 없어 영어 원문으로 표시.
초록 발췌
World-action models (WAMs) jointly predict how a scene will evolve and how an agent should act, however joint generation alone does not necessarily impose a shared geometric constraint on these predictions. We present PhysWAM, a unified world-action model for autonomous driving that co-denoises multiview video, metric depth, and ego motion within a single flow-matching transformer.
초록에서 가져왔습니다. 요약을 준비 중입니다.