Searcharxiv⌕ Search

arXiv subjects

Huaihang Zheng

Publications and source records attributed to Huaihang Zheng.

3 recordsLinked to original sources

TORL-VLA: Tactile Guided Online Reinforcement Learning for Contact-Rich Manipulation

Vision-Language-Action (VLA) models have become a powerful framework for robotic manipulation, and recent studies have introduced tactile or force feedback into VLAs to address contact-rich tasks. However, these models are typically deployed as offline policies. When contact conditions shift from the training distribution, the policy cannot perform online adaptation, leading to problems such as inappropriate contact forces and inefficient retries. Therefore, we propose TORL-VLA, a tactile-guided online reinforcement learning framework that couples tactile feedback with policy refinement for contact-rich manipulation. Our method introduces a tactile-derived wrench-aware VLA to predict reference actions and future wrench sequences, while a lightweight online RL module is used to refine the reference actions. To stabilize learning from mixed exploratory policy-generated and human-intervention data, we introduce an intervention-censored critic that prevents post-intervention success from being wrongly credited to policy-generated actions preceding intervention. Real-robot experiments on long-horizon contact-rich tasks, including latch manipulation, coffee-cup placement, and egg handling, show that TORL-VLA improves success rates at both subtask and full-task levels, as well as time-bounded execution efficiency over strong baselines. Project page: https://torl-vla.github.io/

cs.RO↗

Posture Adjustment for a Wheel-legged Robotic System via Leg Force Control with Prescribed Transient Performance

This work proposes a force control strategy with prescribed transient performance for the legs of a wheel-legged robotic system to realize the posture adjustment on uneven roads. A dynamic model of the robotic system is established with the body postures as inputs and the leg forces as outputs, such that the desired forces for the wheel-legs are calculated by the posture reference and feedback. Based on the funnel control scheme, the legs realize force tracking with prescribed transient performance. To improve the robustness of the force control system, an event-based mechanism is designed for the online segment of the funnel function. As a result, the force tracking error of the wheel-leg evolves inside the performance funnel with proved convergence. The absence of Zeno behavior for the event-triggering condition is also guaranteed. The proposed control scheme is applied to the wheel-legged physical prototype for the performance of force tracking and posture adjustment. Multiple comparative experimental results are presented to validate the stability and effectiveness of the proposed methodology.

cs.RO↗

Quasi Time-Fuel Optimal Control Strategy for dynamic target tracking

This brief proposes a quasi time-fuel optimal control strategy to solve the dynamic tracking problem of unmanned systems when fuel and control input are limited. This kind of motion planning and control strategy could bring the biggest advantage into full play in the field of switching tracking for multiple targets, such as the multi-target strike of weapons and the rapid multi-target grabbing of robots on industrial assembly lines. Compared with the time-fuel optimal control that has been studied before, the proposed controller retains the advantages of optimal control when switching between multiple dynamic targets, and improves the high-frequency oscillation problem of the system under the discontinuous control strategy. Moreover, the asymmetry of friction load, which can affect the dynamic performance of the system, is also considered in this brief. Therefore, the novel control law, which is addressed by this brief, can make the corresponding system achieve the desired transient performance and satisfactory steady-state performance when switching between multiple dynamic targets. The experiment results based on the visual tracking turntable are presented to verify the superiority of the proposed method.

physics.pop-ph↗