Robot Learning with Visual Predicted Force

📅 2026-10-03
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This study addresses the high cost and deployment complexity associated with force-aware manipulation relying on dedicated sensors by proposing a vision-based force estimation method that eliminates the need for tactile or force sensors. Specifically, contact forces are predicted by observing the deformation of a compliant gripper. Furthermore, an action-force joint policy network is constructed to generate candidate actions alongside their expected force feedback, enabling optimal action selection during inference through target-force-guided sampling. This approach facilitates contact-rich robotic manipulation without requiring additional sensing hardware. The effectiveness of the proposed method is validated across practical tasks, including berry harvesting and aluminum can grasping, demonstrating its potential for accessible and robust force-sensitive manipulation in unstructured environments.
📝 Abstract
Force-aware manipulation typically relies on specialized force or tactile sensors. We show that force-aware manipulation can instead be achieved through visual force prediction from the deformation of a compliant Fin Ray gripper. Our approach trains two models. First, we train a visual force estimator on calibration data and use it to annotate task demonstrations with force estimates. Second, we train an action--force proposal policy on these force-augmented demonstrations to jointly generate candidate robot actions and their associated forces. At test time, we sample candidate actions and the forces they are expected to produce, then execute the action whose predicted force is closest to a target from the demonstrations. We evaluate our approach on berry picking, empty-can grasping, in-hand reorientation, and plug insertion. Our results show that visual force prediction can guide inference-time action selection for contact-rich manipulation without requiring force or tactile sensors at deployment.
Problem

Research questions and friction points this paper is trying to address.

force-aware manipulation
visual force prediction
contact-rich manipulation
robot learning
Innovation

Methods, ideas, or system contributions that make the work stand out.

Visual Force Prediction
Force-aware Manipulation
Compliant Gripper
Action-Force Proposal Policy
Contact-rich Manipulation
💼 Related Jobs
No related jobs found.
H
Haonan Chen
Harvard University, Cambridge, MA, USA.
Feiyang Wu
Feiyang Wu
Georgia Institute of Technology
Reinforcement LearningDeep Learning
Y
Yuxiang Ma
Massachusetts Institute of Technology, Cambridge, MA, USA.
M
Mustafa Mete
Harvard University, Cambridge, MA, USA.
P
Pengfei Ye
Massachusetts Institute of Technology, Cambridge, MA, USA.
Junxuan Shen
Junxuan Shen
Massachusetts Institute of Technology, Cambridge, MA, USA.
Cheng Zhu
Cheng Zhu
J. Erskine Love Jr. Endowed Chair in Engineering and Regents' Professor
BiomechanicsMechanobiologyImmunologyCancerHemostasis and Thrombosis
A
Aurora Ruggeri
Harvard University, Cambridge, MA, USA.
K
Kelvin Cheung
Harvard University, Cambridge, MA, USA.
Jiayuan Mao
Jiayuan Mao
MIT CSAIL
Artificial IntelligenceRoboticsComputer VisionNatural Language ProcessingMachine Learning
E
Edward Adelson
Massachusetts Institute of Technology, Cambridge, MA, USA.
Jiajun Wu
Jiajun Wu
Stanford University
Computer VisionRoboticsArtificial IntelligenceMachine LearningCognitive Science
R
Robert D. Howe
Harvard University, Cambridge, MA, USA.
Yilun Du
Yilun Du
Harvard University
Artificial IntelligenceMachine LearningRoboticsComputer Vision