arXiv ScienceSearch

arXiv subjects

Rui Jiang

Publications and source records attributed to Rui Jiang.

At least 37 records · Page 2Linked to original sources

Modeling the residual queue and queue-dependent capacity in a static traffic assignment problem

The residual queue during a given study period (e.g., peak hour) is an important feature that should be considered when solving a traffic assignment problem under equilibrium for strategic traffic planning. Although studies have focused extensively on static or quasi-dynamic traffic assignment models considering the residual queue, they have failed to capture the situation wherein the equilibrium link flow passing through the link is less than the link physical capacity under congested conditions. To address this critical issue, we introduce a novel static traffic assignment model that explicitly incorporates the residual queue and queue-dependent link capacity. The proposed model ensures that equilibrium link flows remain within the physical capacity bounds, yielding estimations more aligned with data observed by traffic detectors, especially in oversaturated scenarios. A generalized link cost function considering queue-dependent capacity, with an additional queuing delay term is proposed. The queuing delay term represents the added travel cost under congestion, offering a framework wherein conventional static models, both with and without physical capacity constraints, become special cases of our model. Our study rigorously analyzes the mathematical properties of the new model, establishing the theoretical uniqueness of solutions for link flow and residual queue under certain conditions. We also introduce a gradient projection-based alternating minimization algorithm tailored for the proposed model. Numerical examples are conducted to demonstrate the superiority and merit of the proposed model and solution algorithm.

eess.SY

Joint Beamforming for Multi-target Detection and Multi-user Communication in ISAC Systems

Detecting weak targets is one of the main challenges for integrated sensing and communication (ISAC) systems. Sensing and communication suffer from a performance trade-off in ISAC systems. As the communication demand increases, sensing ability, especially weak target detection performance, will inevitably reduce. Traditional approaches fail to address this issue. In this paper, we develop a joint beamforming scheme and formulate it as a max-min problem to maximize the detection probability of the weakest target under the constraint of the signal-to-interference-plus-noise ratio (SINR) of multi-user communication. An alternating optimization (AO) algorithm is developed for solving the complicated non-convex problem to obtain the joint beamformer. The proposed scheme can direct the transmit energy toward the multiple targets properly to ensure robust multi-target detection performance. Numerical results show that the proposed beamforming scheme can effectively increase the detection probability of the weakest target compared to baseline approaches while ensuring communication performance.

eess.SP

CamI2V: Camera-Controlled Image-to-Video Diffusion Model

Recent advancements have integrated camera pose as a user-friendly and physics-informed condition in video diffusion models, enabling precise camera control. In this paper, we identify one of the key challenges as effectively modeling noisy cross-frame interactions to enhance geometry consistency and camera controllability. We innovatively associate the quality of a condition with its ability to reduce uncertainty and interpret noisy cross-frame features as a form of noisy condition. Recognizing that noisy conditions provide deterministic information while also introducing randomness and potential misguidance due to added noise, we propose applying epipolar attention to only aggregate features along corresponding epipolar lines, thereby accessing an optimal amount of noisy conditions. Additionally, we address scenarios where epipolar lines disappear, commonly caused by rapid camera movements, dynamic objects, or occlusions, ensuring robust performance in diverse environments. Furthermore, we develop a more robust and reproducible evaluation pipeline to address the inaccuracies and instabilities of existing camera control metrics. Our method achieves a 25.64% improvement in camera controllability on the RealEstate10K dataset without compromising dynamics or generation quality and demonstrates strong generalization to out-of-domain images. Training and inference require only 24GB and 12GB of memory, respectively, for 16-frame sequences at 256x256 resolution. We will release all checkpoints, along with training and evaluation code. Dynamic videos are best viewed at https://zgctroy.github.io/CamI2V.

cs.CV

RSAM-Seg: A SAM-based Approach with Prior Knowledge Integration for Remote Sensing Image Semantic Segmentation

The development of high-resolution remote sensing satellites has provided great convenience for research work related to remote sensing. Segmentation and extraction of specific targets are essential tasks when facing the vast and complex remote sensing images. Recently, the introduction of Segment Anything Model (SAM) provides a universal pre-training model for image segmentation tasks. While the direct application of SAM to remote sensing image segmentation tasks does not yield satisfactory results, we propose RSAM-Seg, which stands for Remote Sensing SAM with Semantic Segmentation, as a tailored modification of SAM for the remote sensing field and eliminates the need for manual intervention to provide prompts. Adapter-Scale, a set of supplementary scaling modules, are proposed in the multi-head attention blocks of the encoder part of SAM. Furthermore, Adapter-Feature are inserted between the Vision Transformer (ViT) blocks. These modules aim to incorporate high-frequency image information and image embedding features to generate image-informed prompts. Experiments are conducted on four distinct remote sensing scenarios, encompassing cloud detection, field monitoring, building detection and road mapping tasks . The experimental results not only showcase the improvement over the original SAM and U-Net across cloud, buildings, fields and roads scenarios, but also highlight the capacity of RSAM-Seg to discern absent areas within the ground truth of certain datasets, affirming its potential as an auxiliary annotation method. In addition, the performance in few-shot scenarios is commendable, underscores its potential in dealing with limited datasets.

cs.CV

Label Informed Contrastive Pretraining for Node Importance Estimation on Knowledge Graphs

Node Importance Estimation (NIE) is a task of inferring importance scores of the nodes in a graph. Due to the availability of richer data and knowledge, recent research interests of NIE have been dedicating to knowledge graphs for predicting future or missing node importance scores. Existing state-of-the-art NIE methods train the model by available labels, and they consider every interested node equally before training. However, the nodes with higher importance often require or receive more attention in real-world scenarios, e.g., people may care more about the movies or webpages with higher importance. To this end, we introduce Label Informed ContrAstive Pretraining (LICAP) to the NIE problem for being better aware of the nodes with high importance scores. Specifically, LICAP is a novel type of contrastive learning framework that aims to fully utilize the continuous labels to generate contrastive samples for pretraining embeddings. Considering the NIE problem, LICAP adopts a novel sampling strategy called top nodes preferred hierarchical sampling to first group all interested nodes into a top bin and a non-top bin based on node importance scores, and then divide the nodes within top bin into several finer bins also based on the scores. The contrastive samples are generated from those bins, and are then used to pretrain node embeddings of knowledge graphs via a newly proposed Predicate-aware Graph Attention Networks (PreGAT), so as to better separate the top nodes from non-top nodes, and distinguish the top nodes within top bin by keeping the relative order among finer bins. Extensive experiments demonstrate that the LICAP pretrained embeddings can further boost the performance of existing NIE methods and achieve the new state-of-the-art performance regarding both regression and ranking metrics. The source code for reproducibility is available at https://github.com/zhangtia16/LICAP

cs.AI

Adaptive Kalman-based hybrid car following strategy using TD3 and CACC

In autonomous driving, the hybrid strategy of deep reinforcement learning and cooperative adaptive cruise control (CACC) can fully utilize the advantages of the two algorithms and significantly improve the performance of car following. However, it is challenging for the traditional hybrid strategy based on fixed coefficients to adapt to mixed traffic flow scenarios, which may decrease the performance and even lead to accidents. To address the above problems, a hybrid car following strategy based on an adaptive Kalman Filter is proposed by regarding CACC and Twin Delayed Deep Deterministic Policy Gradient (TD3) algorithms. Different from traditional hybrid strategy based on fixed coefficients, the Kalman gain H, using as an adaptive coefficient, is derived from multi-timestep predictions and Monte Carlo Tree Search. At the end of study, simulation results with 4157745 timesteps indicate that, compared with the TD3 and HCFS algorithms, the proposed algorithm in this study can substantially enhance the safety of car following in mixed traffic flow without compromising the comfort and efficiency.

cs.AI

Quasinormal modes of the spherical bumblebee black holes with a global monopole

The bumblebee model is an extension of the Einstein-Maxwell theory that allows for the spontaneous breaking of the Lorentz symmetry of the spacetime. In this paper, we study the quasinormal modes of the spherical black holes in this model that are characterized by a global monopole. We analyze the two cases with a vanishing cosmological constant or a negative one (the anti-de Sitter case). We find that the black holes are stable under the perturbation of a massless scalar field. However, both the Lorentz symmetry breaking and the global monopole have notable impacts on the evolution of the perturbation. The Lorentz symmetry breaking may prolong or shorten the decay of the perturbation according to the sign of the breaking parameter. The global monopole, on the other hand, has different effects depending on whether a nonzero cosmological constant presences: it reduces the damping of the perturbations for the case with a vanishing cosmological constant, but has little influence for the anti-de Sitter case.

gr-qc

One-stop Training of Multiple Capacity Models

Training models with varying capacities can be advantageous for deploying them in different scenarios. While high-capacity models offer better performance, low-capacity models require fewer computing resources for training and inference. In this work, we propose a novel one-stop training framework to jointly train high-capacity and low-capactiy models. This framework consists of two composite model architectures and a joint training algorithm called Two-Stage Joint-Training (TSJT). Unlike knowledge distillation, where multiple capacity models are trained from scratch separately, our approach integrates supervisions from different capacity models simultaneously, leading to faster and more efficient convergence. Extensive experiments on the multilingual machine translation benchmark WMT10 show that our method outperforms low-capacity baseline models and achieves comparable or better performance on high-capacity models. Notably, the analysis demonstrates that our method significantly influences the initial training process, leading to more efficient convergence and superior solutions.

cs.CL

Experimental features of emissions and fuel consumption in a car-following platoon

The paper investigates the features of emissions and fuel consumption (EFC) in a car-following (CF) platoon based on two experimental datasets. Four classical EFC models are employed and a universal concave growth pattern of the EFC along a platoon has been demonstrated. A general framework of coupling EFC and CF models is tested by calibrating and simulating three classical CF models. This work first demonstrates that, at vehicle-pair level, all models perform well on EFC prediction. The intelligent driver model outperforms the other CF models on calibration accuracy, but this is not true on EFC prediction. Second, at platoon level, the predicted EFC is nearly constant along the platoon which qualitatively differs from the experimental observation. The investigation highlights that accurate estimations at vehicle level may be insufficient for analysis at platoon level due to the significant role of oscillation growth and evolution in EFC estimation.

physics.soc-ph

Time-aware Multiway Adaptive Fusion Network for Temporal Knowledge Graph Question Answering

Knowledge graphs (KGs) have received increasing attention due to its wide applications on natural language processing. However, its use case on temporal question answering (QA) has not been well-explored. Most of existing methods are developed based on pre-trained language models, which might not be capable to learn \emph{temporal-specific} presentations of entities in terms of temporal KGQA task. To alleviate this problem, we propose a novel \textbf{T}ime-aware \textbf{M}ultiway \textbf{A}daptive (\textbf{TMA}) fusion network. Inspired by the step-by-step reasoning behavior of humans. For each given question, TMA first extracts the relevant concepts from the KG, and then feeds them into a multiway adaptive module to produce a \emph{temporal-specific} representation of the question. This representation can be incorporated with the pre-trained KG embedding to generate the final prediction. Empirical results verify that the proposed model achieves better performance than the state-of-the-art models in the benchmark dataset. Notably, the Hits@1 and Hits@10 results of TMA on the CronQuestions dataset's complex questions are absolutely improved by 24\% and 10\% compared to the best-performing baseline. Furthermore, we also show that TMA employing an adaptive fusion mechanism can provide interpretability by analyzing the proportion of information in question representations.

cs.CL

On the calibration of stochastic car following models

Recent experimental and empirical observations have demonstrated that stochasticity plays a critical role in car following (CF) dynamics. To reproduce the observations, quite a few stochastic CF models have been proposed. However, while calibrating the deterministic CF models is well investigated, studies on how to calibrate the stochastic models are lacking. Motivated by this fact, this paper aims to address this fundamental research gap. Firstly, the CF experiment under the same driving environment is conducted and analyzed. Based on the experimental results, we test two previous calibration methods, i.e., the method to minimize the Multiple Runs Mean (MRMean) error and the method of maximum likelihood estimation (MLE). Deficiencies of the two methods have been identified. Next, we propose a new method to minimize the Multiple Runs Minimum (MRMin) error. Calibration based on the experimental data and the synthetic data demonstrates that the new method outperforms the two previous methods. Furthermore, the mechanisms of different methods are explored from the perspective of error analysis. The analysis indicates that the new method can be regarded as a nested optimization model. The method separates the aleatoric errors caused by stochasticity from the epistemic error caused by parameters, and it is able to deal with the two kinds of errors effectively. Finally, we find that under the calibration framework of stochastic CF models, the calibrated parameter set using spacing as MoP may not always outperform that using velocity as MoP. These findings are expected to enhance the understanding of the role of stochasticity in CF dynamics where the new calibration framework for stochastic CF models is established.

physics.soc-ph

Experimental study and modeling of the lower-level controller of automated vehicle

Accurate modeling of lower-level controller plays an important role in the traffic flow of automated vehicles (AVs). However, there lacks enough attention with this respect. To address this issue, we conduct a field experiment with two vehicles that are equipped with developable autonomous driving system, where one can customize the upper-level control algorithm. Based on the field experimental data, a new lower-level control model is developed and compared with two widely used ones. The comparison results show that the proposed model outperforms the two previous models in capturing the observed actual acceleration, especially the troughs of the acceleration time series. Furthermore, theoretical analysis indicates that comparing with the proposed model, the two previous models significantly overestimate the stability region of the traffic flow of the AVs and the capacity of stable traffic flow. Our study is expected to further shed light on the importance of accurate lower-level control modeling.

physics.soc-ph

ROSE: Robust Selective Fine-tuning for Pre-trained Language Models

Even though the large-scale language models have achieved excellent performances, they suffer from various adversarial attacks. A large body of defense methods has been proposed. However, they are still limited due to redundant attack search spaces and the inability to defend against various types of attacks. In this work, we present a novel fine-tuning approach called \textbf{RO}bust \textbf{SE}letive fine-tuning (\textbf{ROSE}) to address this issue. ROSE conducts selective updates when adapting pre-trained models to downstream tasks, filtering out invaluable and unrobust updates of parameters. Specifically, we propose two strategies: the first-order and second-order ROSE for selecting target robust parameters. The experimental results show that ROSE achieves significant improvements in adversarial robustness on various downstream NLP tasks, and the ensemble method even surpasses both variants above. Furthermore, ROSE can be easily incorporated into existing fine-tuning methods to improve their adversarial robustness further. The empirical analysis confirms that ROSE eliminates unrobust spurious updates during fine-tuning, leading to solutions corresponding to flatter and wider optima than the conventional method. Code is available at \url{https://github.com/jiangllan/ROSE}.

cs.CL

DSLA: Dynamic smooth label assignment for efficient anchor-free object detection

Anchor-free detectors basically formulate object detection as dense classification and regression. For popular anchor-free detectors, it is common to introduce an individual prediction branch to estimate the quality of localization. The following inconsistencies are observed when we delve into the practices of classification and quality estimation. Firstly, for some adjacent samples which are assigned completely different labels, the trained model would produce similar classification scores. This violates the training objective and leads to performance degradation. Secondly, it is found that detected bounding boxes with higher confidences contrarily have smaller overlaps with the corresponding ground-truth. Accurately localized bounding boxes would be suppressed by less accurate ones in the Non-Maximum Suppression (NMS) procedure. To address the inconsistency problems, the Dynamic Smooth Label Assignment (DSLA) method is proposed. Based on the concept of centerness originally developed in FCOS, a smooth assignment strategy is proposed. The label is smoothed to a continuous value in [0, 1] to make a steady transition between positive and negative samples. Intersection-of-Union (IoU) is predicted dynamically during training and is coupled with the smoothed label. The dynamic smooth label is assigned to supervise the classification branch. Under such supervision, quality estimation branch is naturally merged into the classification branch, which simplifies the architecture of anchor-free detector. Comprehensive experiments are conducted on the MS COCO benchmark. It is demonstrated that, DSLA can significantly boost the detection accuracy by alleviating the above inconsistencies for anchor-free detectors. Our codes are released at https://github.com/YonghaoHe/DSLA.

cs.CV

Stability analysis of stochastic second-order macroscopic continuum models and numerical simulations

Second-order macroscopic continuum models have been constantly improving for decades to reproduce the empirical observations. Recently, a series of experimental studies have suggested that the stochastic factors contribute significantly to destabilizing traffic flow. Nevertheless, the traffic flow stability of the stochastic second-order macroscopic continuum model hasn't received the attention it deserves in past studies. More importantly, we have found that the destabilizing aspect of stochasticity is still not correctly validated in the existing theoretical stability analysis. In this paper, we analytically study the impact of stochasticity on traffic flow stability for a general stochastic second-order macroscopic model by using the direct Lyapunov method. Numerical simulations have been carried out for different typical stochastic second-order macroscopic models. Our analytical stability analysis has been validated, and our methodology has been proved more efficient. Our study has theoretically revealed that the presence of stochasticity has a destabilizing effect in stochastic macroscopic models.

physics.soc-ph

Oscillation growth in mixed traffic flow of human driven vehicles and automated vehicles: Experimental study and simulation

This paper reports an experimental study on oscillation growth in mixed traffic flow of automated vehicles (AVs) and human driven vehicles (HVs). The leading vehicle moves with constant speed in the experiment. The following vehicles consist of six developable AVs and different number of HVs. Thus, the market penetration rate (MPR) of AVs decreases with the increase of platoon size. The AVs are homogeneously distributed in the platoon. The constant time gap car-following policy is adopted for the AVs and the gap is set to 1.5 s. The experiment shows that in the 7-vehicle-platoon, the oscillations grow only slightly. In the 10-vehicle-platoon, the AVs could still significantly suppress the growth of oscillations. With the further decrease of MPR of AVs in the 13- and 20-vehicle-platoon, the AVs become having no significant impact on oscillation growth. On the other hand, with the decrease of MPR of AVs, average density of the vehicles and flow rate of the platoon increase, which demonstrates a trade-off between traffic stability and throughput under the given setup of AVs. The simulation study is also carried out, which exhibits good agreement with the experiment. Finally, sensitivity analysis of the parameters in the AV upper-level control algorithm has been performed, which is expected to guide future experiment design.

physics.soc-ph

Stochastic factors and string stability of traffic flow: Analytical investigation and numerical study based on car-following models

The emergence dynamics of traffic instability has always attracted particular attention. For several decades, researchers have studied the stability of traffic flow using deterministic traffic models, with less emphasis on the presence of stochastic factors. However, recent empirical and theoretical findings have demonstrated that the stochastic factors tend to destabilize traffic flow and stimulate the concave growth pattern of traffic oscillations. In this paper, we derive a string stability condition of a general stochastic continuous car-following model by the mean of the generalized Lyapunov equation. We have found, indeed, that the presence of stochasticity destabilizes the traffic flow. The impact of stochasticity depends on both the sensitivity to the gap and the sensitivity to the velocity difference. Numerical simulations of three typical car-following models have been carried out to validate our theoretical analysis. Finally, we have calibrated and validated the stochastic car-following models against empirical data. It is found that the stochastic car-following models reproduce the observed traffic instability and capture the concave growth pattern of traffic oscillations. Our results further highlight theoretically and numerically that the stochastic factors have a significant impact on traffic dynamics.

physics.soc-ph

Unpaired Quad-Path Cycle Consistent Adversarial Networks for Single Image Defogging

Adversarial learning-based image defogging methods have been extensively studied in computer vision due to their remarkable performance. However, most existing methods have limited defogging capabilities for real cases because they are trained on the paired clear and synthesized foggy images of the same scenes. In addition, they have limitations in preserving vivid color and rich textual details in defogging. To address these issues, we develop a novel generative adversarial network, called quad-path cycle consistent adversarial network (QPC-Net), for single image defogging. QPC-Net consists of a Fog2Fogfree block and a Fogfree2Fog block. In each block, there are three learning-based modules, namely, fog removal, color-texture recovery, and fog synthetic, which sequentially compose dual-path that constrain each other to generate high quality images. Specifically, the color-texture recovery model is designed to exploit the self-similarity of texture and structure information by learning the holistic channel-spatial feature correlations between the foggy image with its several derived images. Moreover, in the fog synthetic module, we utilize the atmospheric scattering model to guide it to improve the generative quality by focusing on an atmospheric light optimization with a novel sky segmentation network. Extensive experiments on both synthetic and real-world datasets show that QPC-Net outperforms state-of-the-art defogging methods in terms of quantitative accuracy and subjective visual quality.

cs.CV