arXiv ScienceSearch

arXiv subjects

Haohan Yang

Publications and source records attributed to Haohan Yang.

3 recordsLinked to original sources

Reinforced Refinement with Self-Aware Expansion for End-to-End Autonomous Driving

End-to-end autonomous driving has emerged as a promising paradigm for directly mapping sensor inputs to planning maneuvers using learning-based modular integrations. However, existing imitation learning (IL)-based models suffer from generalization to hard cases, and a lack of corrective feedback loop under post-deployment. While reinforcement learning (RL) offers a potential solution to tackle hard cases with optimality, it is often hindered by overfitting to specific driving cases, resulting in catastrophic forgetting of generalizable knowledge and sample inefficiency. To overcome these challenges, we propose Reinforced Refinement with Self-aware Expansion (R2SE), a novel learning pipeline that constantly refines hard domain while keeping generalizable driving policy for model-agnostic end-to-end driving systems. Through reinforcement fine-tuning and policy expansion that facilitates continuous improvement, R2SE features three key components: 1) Generalist Pretraining with hard-case allocation trains a generalist imitation learning (IL) driving system while dynamically identifying failure-prone cases for targeted refinement; 2) Residual Reinforced Specialist Fine-tuning optimizes residual corrections using reinforcement learning (RL) to improve performance in hard case domain while preserving global driving knowledge; 3) Self-aware Adapter Expansion dynamically integrates specialist policies back into the generalist model, enhancing continuous performance improvement. Experimental results in closed-loop simulation and real-world datasets demonstrate improvements in generalization, safety, and long-horizon policy robustness over state-of-the-art E2E systems, highlighting the effectiveness of reinforce refinement for scalable autonomous driving.

cs.RO

Hybrid-Prediction Integrated Planning for Autonomous Driving

Autonomous driving systems require the ability to fully understand and predict the surrounding environment to make informed decisions in complex scenarios. Recent advancements in learning-based systems have highlighted the importance of integrating prediction and planning modules. However, this integration has brought forth three major challenges: inherent trade-offs by sole prediction, consistency between prediction patterns, and social coherence in prediction and planning. To address these challenges, we introduce a hybrid-prediction integrated planning (HPP) system, which possesses three novelly designed modules. First, we introduce marginal-conditioned occupancy prediction to align joint occupancy with agent-wise perceptions. Our proposed MS-OccFormer module achieves multi-stage alignment per occupancy forecasting with consistent awareness from agent-wise motion predictions. Second, we propose a game-theoretic motion predictor, GTFormer, to model the interactive future among individual agents with their joint predictive awareness. Third, hybrid prediction patterns are concurrently integrated with Ego Planner and optimized by prediction guidance. HPP achieves state-of-the-art performance on the nuScenes dataset, demonstrating superior accuracy and consistency for end-to-end paradigms in prediction and planning. Moreover, we test the long-term open-loop and closed-loop performance of HPP on the Waymo Open Motion Dataset and CARLA benchmark, surpassing other integrated prediction and planning pipelines with enhanced accuracy and compatibility.

cs.RO

A PXI-based Multi-channel Data Acquisition System for Fast Transient Pulses

In this paper, we design a PXI-based, multi-channel data-acquisition system (DAS) mainly applicable to recording one-shot fast transient pulses in nuclear physics experiments. The system consists of one NI PXIe-1085 chassis, containing a controller card and at most 16 data-acquisition (DAQ) cards. Every single DAQ card has a sampling rate of 1GS/s and a 12bit vertical resolution with the PXI interface and SFP+ transceiver for data transmission. When the system is put into operation near the pulsed radiation source, the SFP+ optical fiber channel enables a timely data transmission to a remote server. All of these cards in the chassis can be synchronized using PXI timing and triggering resources. Additionally, a simple DAS software is developed to display the pulsed signals captured and communicate with the host PC for remote control and data upload. After careful calibration, preliminary tests show that every DAQ channel achieves an analog bandwidth higher than 200MHz and an ENOB of more than 9 bits at a 1GS/s sampling rate. Owing to such high speed and resolution, the system may facilitate improvements in extracting maximum information from transient signals. Furthermore, with great scalability and high-speed data transmission, the system can be used for other nuclear physics experiments.

physics.ins-det