Is this you? As a journalist, you can create a free Muck Rack account to customize your profile, list your contact preferences, and upload a portfolio of your best work.
Claim your profile
Get in touch with Siyin
Contact Siyin, search articles and posts on X, monitor coverage, and track replies from one place.
Learn more about Muck RackActions
Is this you?
As a journalist, you can create a free Muck Rack account to customize your profile, list your contact preferences, and upload a portfolio of your best work.Articles
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning
Abstract Reinforcement learning (RL) post-training of Vision-Language-Action (VLA) models has shown strong promise for robotic manipulation. Among RL methods, critic-based approaches rely on a value estimator that predominantly operates on single-frame observations or single-frame VLM backbone latents, which is a fundamental mismatch with the partially observable nature of robot control.
Paper page - World Modeling Makes a Better Planner: Dual Preference Optimization for Embodied Task Planning
📣 World Modeling Makes a Better Planner: Dual Preference Optimization for Embodied Task Planning 🤔 Current LVLMs struggle with grounding in embodied environments, how can we make AI agents understand the physical world like humans?
Actions
Is this you?
As a journalist, you can create a free Muck Rack account to customize your profile, list your contact preferences, and upload a portfolio of your best work.Get in touch with Siyin
Contact Siyin, search articles and posts on X, monitor coverage, and track replies from one place.
Learn more about Muck Rack