arXiv Science⌕ Search

arXiv subjects

Fu Li

Publications and source records attributed to Fu Li.

At least 91 records · Page 5Linked to original sources

NTIRE 2020 Challenge on Video Quality Mapping: Methods and Results

This paper reviews the NTIRE 2020 challenge on video quality mapping (VQM), which addresses the issues of quality mapping from source video domain to target video domain. The challenge includes both a supervised track (track 1) and a weakly-supervised track (track 2) for two benchmark datasets. In particular, track 1 offers a new Internet video benchmark, requiring algorithms to learn the map from more compressed videos to less compressed videos in a supervised training manner. In track 2, algorithms are required to learn the quality mapping from one device to another when their quality varies substantially and weakly-aligned video pairs are available. For track 1, in total 7 teams competed in the final test phase, demonstrating novel and effective solutions to the problem. For track 2, some existing methods are evaluated, showing promising solutions to the weakly-supervised video quality mapping problem.

eess.IV↗

Single Crystalline Colloidal Quasi-Two-Dimensional Tin Telluride

Tin telluride is a narrow gap semiconductor with promising properties for IR optical applications and topological insulators. We report a convenient colloidal synthesis of quasi-two-dimensional SnTe nanocrystals through the hot-injection method in a non-polar solvent. By introducing the halide alkane 1-bromotetradecane as well as oleic acid and trioctylphosphine, the thickness of two dimensional SnTe nanostripes can be tuned down to 30 nm, while the lateral dimensional can reach 6 microns. The obtained SnTe nanostripes are single-crystalline with a rock-salt crystal structure. The absorption spectra demonstrate pronounced absorption features in the IR range revealing the effect of quantum confinement in such structures.

cond-mat.mtrl-sci↗

NTIRE 2020 Challenge on Perceptual Extreme Super-Resolution: Methods and Results

This paper reviews the NTIRE 2020 challenge on perceptual extreme super-resolution with focus on proposed solutions and results. The challenge task was to super-resolve an input image with a magnification factor 16 based on a set of prior examples of low and corresponding high resolution images. The goal is to obtain a network design capable to produce high resolution results with the best perceptual quality and similar to the ground truth. The track had 280 registered participants, and 19 teams submitted the final results. They gauge the state-of-the-art in single image super-resolution.

eess.IV↗

Temporal Quantum Noise Reduction Acquired by an Electron-Multiplying Charge-Coupled-Device Camera

Electron-multiplying charge-coupled-device cameras (EMCCDs) have been used to observe quantum noise reductions in beams of light in the transverse spatial degree of freedom. For the quantum noise reduction in the temporal domain, "bucket detectors," usually composed of photodiodes with operational amplifiers, are used to register the intensity fluctuations in beams of light within the bandwidth of the detectors. Spatial information, however, is inevitably washed off by the detector. In this paper, we report on measurements of the temporal quantum noise reduction in bright twin beams using an EMCCD camera. The four-wave mixing process in an atomic rubidium vapor cell is used to generate the bright twin beams of light. We observe more than 25% of temporal quantum noise reduction with respect to the shot-noise limit in images captured by the EMCCD camera. Compared with bucket detectors, EMCCD makes it possible to take advantage of the spatial and temporal quantum properties of light simultaneously, which would greatly benefit many applications using quantum protocols.

physics.optics↗

Cost-effectiveness Analysis of Antiepidemic Policies and Global Situation Assessment of COVID-19

With a two-layer contact-dispersion model and data in China, we analyze the cost-effectiveness of three types of antiepidemic measures for COVID-19: regular epidemiological control, local social interaction control, and inter-city travel restriction. We find that: 1) intercity travel restriction has minimal or even negative effect compared to the other two at the national level; 2) the time of reaching turning point is independent of the current number of cases, and only related to the enforcement stringency of epidemiological control and social interaction control measures; 3) strong enforcement at the early stage is the only opportunity to maximize both antiepidemic effectiveness and cost-effectiveness; 4) mediocre stringency of social interaction measures is the worst choice. Subsequently, we cluster countries/regions into four groups based on their control measures and provide situation assessment and policy suggestions for each group.

physics.soc-ph↗

Anisotropic circular photogalvanic effect in colloidal tin sulfide nanosheets

Tin sulfide promises very interesting properties such as a high optical absorption coefficient and a small band gap, while being less toxic compared to other metal chalcogenides. However, the limitations in growing atomically thin structures of tin sulfide hinder the experimental realization of these properties. Due to the flexibility of the colloidal synthesis, it is possible to synthesize very thin and at the same time large nanosheets. Electrical transport measurements show that these nanosheets can function as field-effect transistors with high on/off ratio and p-type behavior. The temperature dependency of the charge transport reveals that defects in the crystal are responsible for the formation of holes as majority carriers. During illumination with circularly polarized light, these crystals generate a helicity dependent photocurrent at zero-volt bias, since their symmetry is broken by asymmetric interfaces (substrate and vacuum). Further, the observed circular photogalvanic effect shows a pronounced in-plane anisotropy, with a higher photocurrent along the armchair direction, originating from the higher absorption coefficient in this direction. Our new insights show the potential of tin sulfide for new functionalities in electronics and optoelectronics, for instance as polarization sensors.

cond-mat.mtrl-sci↗

Multi-Label Classification with Label Graph Superimposing

Images or videos always contain multiple objects or actions. Multi-label recognition has been witnessed to achieve pretty performance attribute to the rapid development of deep learning technologies. Recently, graph convolution network (GCN) is leveraged to boost the performance of multi-label recognition. However, what is the best way for label correlation modeling and how feature learning can be improved with label system awareness are still unclear. In this paper, we propose a label graph superimposing framework to improve the conventional GCN+CNN framework developed for multi-label recognition in the following two aspects. Firstly, we model the label correlations by superimposing label graph built from statistical co-occurrence information into the graph constructed from knowledge priors of labels, and then multi-layer graph convolutions are applied on the final superimposed graph for label embedding abstraction. Secondly, we propose to leverage embedding of the whole label system for better representation learning. In detail, lateral connections between GCN and CNN are added at shallow, middle and deep layers to inject information of label system into backbone CNN for label-awareness in the feature learning process. Extensive experiments are carried out on MS-COCO and Charades datasets, showing that our proposed solution can greatly improve the recognition performance and achieves new state-of-the-art recognition performance.

cs.CV↗

Photon statistics of quantum light on scattering from rotating ground glass

When a laser beam passes through a rotating ground glass (RGG), the scattered light exhibits thermal statistics. This is extensively used in speckle imaging. This scattering process has not been addressed in photon picture and is especially relevant if non-classical light is scattered by the RGG. We develop the photon picture for the scattering process using the Bose statistics for distributing $N$ photons in $M$ pixels. We obtain analytical form for the P-distribution of the output field in terms of the P-distribution of the input field. In particular we obtain a general relation for the $n$-th order correlation function of the scattered light, i.e., $g_{\text{out}}^{(n)}\simeq n!\,g_{\text{in}}^{(n)}$, which holds for any order-$n$ and for arbitrary input states. This result immediately recovers the classical transformation of coherent light to pseudo-thermal light by RGG.

quant-ph↗

TruNet: Short Videos Generation from Long Videos via Story-Preserving Truncation

In this work, we introduce a new problem, named as {\em story-preserving long video truncation}, that requires an algorithm to automatically truncate a long-duration video into multiple short and attractive sub-videos with each one containing an unbroken story. This differs from traditional video highlight detection or video summarization problems in that each sub-video is required to maintain a coherent and integral story, which is becoming particularly important for resource-production video sharing platforms such as Youtube, Facebook, TikTok, Kwai, etc. To address the problem, we collect and annotate a new large video truncation dataset, named as TruNet, which contains 1470 videos with on average 11 short stories per video. With the new dataset, we further develop and train a neural architecture for video truncation that consists of two components: a Boundary Aware Network (BAN) and a Fast-Forward Long Short-Term Memory (FF-LSTM). We first use the BAN to generate high quality temporal proposals by jointly considering frame-level attractiveness and boundaryness. We then apply the FF-LSTM, which tends to capture high-order dependencies among a sequence of frames, to decide whether a temporal proposal is a coherent and integral story. We show that our proposed framework outperforms existing approaches for the story-preserving long video truncation problem in both quantitative measures and user-study. The dataset is available for public academic research usage at https://ai.baidu.com/broad/download.

cs.CV↗

Deep Concept-wise Temporal Convolutional Networks for Action Localization

Existing action localization approaches adopt shallow temporal convolutional networks (\ie, TCN) on 1D feature map extracted from video frames. In this paper, we empirically find that stacking more conventional temporal convolution layers actually deteriorates action classification performance, possibly ascribing to that all channels of 1D feature map, which generally are highly abstract and can be regarded as latent concepts, are excessively recombined in temporal convolution. To address this issue, we introduce a novel concept-wise temporal convolution (CTC) layer as an alternative to conventional temporal convolution layer for training deeper action localization networks. Instead of recombining latent concepts, CTC layer deploys a number of temporal filters to each concept separately with shared filter parameters across concepts. Thus can capture common temporal patterns of different concepts and significantly enrich representation ability. Via stacking CTC layers, we proposed a deep concept-wise temporal convolutional network (C-TCN), which boosts the state-of-the-art action localization performance on THUMOS'14 from 42.8 to 52.1 in terms of mAP(\%), achieving a relative improvement of 21.7\%. Favorable result is also obtained on ActivityNet.

cs.CV↗

Beyond sub-Rayleigh imaging via high order correlation of speckle illumination

Second order intensity correlations of speckle illumination are extensively used in imaging applications that require going beyond the Rayleigh limit. The theoretical analysis shows that significantly improved imaging can be extracted from the study of increasingly higher order intensity cumulants. We provide experimental evidence by demonstrating resolution beyond what is achievable by second order correlations. We present results up to 20th order. We also show an increased visibility of cumulant correlations compared to moment correlations. Our findings clearly suggest the benefits of using higher order intensity cumulants in other disciplines like astronomy and biology.

physics.optics↗

In-plane anisotropic faceting of ultralarge and thin single-crystalline colloidal SnS nanosheets

The colloidal synthesis of large thin two-dimensional (2D) nanosheets is fascinating but challenging, since the growth along the lateral and vertical dimensions need to be controlled independently. In-plane anisotropy in 2D nanosheets is attracting more attention as well. We present a new synthesis for large colloidal single-crystalline SnS nanosheets with the thicknesses down to 7 nm and lateral sizes up to 8 um. The synthesis uses trioctylphosphine-S (TOP-S) as sulfur source and oleic acid (with or without TOP) as ligands. Upon adjusting the capping ligand amount, the growth direction can be switched between anisotropic directions (armchair and zigzag) and isotropic directions ("ladder" directions), leading to an edge-morphology anisotropy. This is the first report on solution-phase synthesis of large thin SnS NSs with tunable edge faceting. Furthermore, electronic transport measurements show strong dependency on the crystallographic directions confirming structural anisotropy.

cond-mat.mtrl-sci↗

Read, Watch, and Move: Reinforcement Learning for Temporally Grounding Natural Language Descriptions in Videos

The task of video grounding, which temporally localizes a natural language description in a video, plays an important role in understanding videos. Existing studies have adopted strategies of sliding window over the entire video or exhaustively ranking all possible clip-sentence pairs in a pre-segmented video, which inevitably suffer from exhaustively enumerated candidates. To alleviate this problem, we formulate this task as a problem of sequential decision making by learning an agent which regulates the temporal grounding boundaries progressively based on its policy. Specifically, we propose a reinforcement learning based framework improved by multi-task learning and it shows steady performance gains by considering additional supervised boundary information during training. Our proposed framework achieves state-of-the-art performance on ActivityNet'18 DenseCaption dataset and Charades-STA dataset while observing only 10 or less clips per video.

cs.CV↗

StNet: Local and Global Spatial-Temporal Modeling for Action Recognition

Despite the success of deep learning for static image understanding, it remains unclear what are the most effective network architectures for the spatial-temporal modeling in videos. In this paper, in contrast to the existing CNN+RNN or pure 3D convolution based approaches, we explore a novel spatial temporal network (StNet) architecture for both local and global spatial-temporal modeling in videos. Particularly, StNet stacks N successive video frames into a \emph{super-image} which has 3N channels and applies 2D convolution on super-images to capture local spatial-temporal relationship. To model global spatial-temporal relationship, we apply temporal convolution on the local spatial-temporal feature maps. Specifically, a novel temporal Xception block is proposed in StNet. It employs a separate channel-wise and temporal-wise convolution over the feature sequence of video. Extensive experiments on the Kinetics dataset demonstrate that our framework outperforms several state-of-the-art approaches in action recognition and can strike a satisfying trade-off between recognition accuracy and model complexity. We further demonstrate the generalization performance of the leaned video representations on the UCF101 dataset.

cs.CV↗

Colloidal Tin Sulfide Nanosheets: Formation Mechanism, Ligand-mediated Shape Tuning and Photo-detection

Colloidal materials of tin(II) sulfide (SnS), as a layered semiconductor with a narrow band gap, are emerging as a potential alternative to the more toxic metal chalcogenides (PbS, PbSe, CdS, CdSe) for various applications such as electronic and optoelectronic devices. We describe a new and simple pathway to produce colloidal SnS nanosheets with large lateral sizes and controllable thickness, as well as single-crystallinity. The synthesis of the nanosheets is achieved by employing tin(II) acetate as tin precursor instead of harmful precursors such as bis[bis(trimethylsilyl)amino] tin(II) and halogen-involved precursors like tin chloride, which limits the large-scale production. We successfully tuned the morphology between squared nanosheets with lateral dimensions from 150 to about 500 nm and a thickness from 24 to 29 nm, and hexagonal nanosheets with lateral sizes from 230 to 1680 nm and heights ranging from 16 to 50 nm by varying the ligands oleic acid and trioctylphosphine. The formation mechanism of both shapes has been investigated in depth, which is also supported by DFT simulations. The optoelectronic measurements show their relatively high conductivity with a pronounced sensitivity to light, which is promising in terms of photo-switching, photo-sensing, and photovoltaic applications also due to their reduced toxicity.

cond-mat.mtrl-sci↗

Combinatorial Multi-Armed Bandit with General Reward Functions

In this paper, we study the stochastic combinatorial multi-armed bandit (CMAB) framework that allows a general nonlinear reward function, whose expected value may not depend only on the means of the input random variables but possibly on the entire distributions of these variables. Our framework enables a much larger class of reward functions such as the $\max()$ function and nonlinear utility functions. Existing techniques relying on accurate estimations of the means of random variables, such as the upper confidence bound (UCB) technique, do not work directly on these functions. We propose a new algorithm called stochastically dominant confidence bound (SDCB), which estimates the distributions of underlying random variables and their stochastically dominant confidence bounds. We prove that SDCB can achieve $O(\log{T})$ distribution-dependent regret and $\tilde{O}(\sqrt{T})$ distribution-independent regret, where $T$ is the time horizon. We apply our results to the $K$-MAX problem and expected utility maximization problems. In particular, for $K$-MAX, we provide the first polynomial-time approximation scheme (PTAS) for its offline problem, and give the first $\tilde{O}(\sqrt T)$ bound on the $(1-ε)$-approximation regret of its online problem, for any $ε>0$.

cs.LG↗

Exploiting Spatial-Temporal Modelling and Multi-Modal Fusion for Human Action Recognition

In this report, our approach to tackling the task of ActivityNet 2018 Kinetics-600 challenge is described in detail. Though spatial-temporal modelling methods, which adopt either such end-to-end framework as I3D \cite{i3d} or two-stage frameworks (i.e., CNN+RNN), have been proposed in existing state-of-the-arts for this task, video modelling is far from being well solved. In this challenge, we propose spatial-temporal network (StNet) for better joint spatial-temporal modelling and comprehensively video understanding. Besides, given that multi-modal information is contained in video source, we manage to integrate both early-fusion and later-fusion strategy of multi-modal information via our proposed improved temporal Xception network (iTXN) for video understanding. Our StNet RGB single model achieves 78.99\% top-1 precision in the Kinetics-600 validation set and that of our improved temporal Xception network which integrates RGB, flow and audio modalities is up to 82.35\%. After model ensemble, we achieve top-1 precision as high as 85.0\% on the validation set and rank No.1 among all submissions.

cs.CV↗

Investigating the `Past of a Particle' without disturbing it

In a recent article [Chin. Phys. Lett. 34, 020301 (2017)], Ben-Israel et al. have claimed that the experiment proposed in [Chin. Phys. Lett. 32, 050303 (2015)] to determine the past of a quantum particle in a nested Mach-Zehnder interferometer does not work, and they have proposed a modification to the experiment. We show that their claim is false, and the modification is not required.

quant-ph↗