Explore, Execute, Evolve: A Skill Acquisition and Reuse Loop for Embodied Agents
From the abstract
Vision-language-action and world-action models have demonstrated impressive capabilities in robotics, yet generalization to unseen tasks remains challenging. More recently, general-purpose multimodal agents have shown great potential for zero-shot robotic task solving.
From the abstract. Our summary is in progress.