arXiv ScienceSearch

arXiv subjects

Eric Marchand

Publications and source records attributed to Eric Marchand.

14 recordsLinked to original sources

Unbiased estimation in one-parameter exponential families for the inverse of the natural parameter with extensions

For one-parameter continuous exponential families, we identify an unbiased estimator of the inverse of the natural parameter $\theta$ for cases where $\theta > 0$, extending an earlier result of \cite{voinov1985unbiased} applicable to a normal model. We provide various applications for Gamma models, Inverse Gaussian models, distributions obtained by truncation, and ratios of normal means. Moreover, we extend the findings to estimating negative powers $\theta^{-k}$, and more generally to complete monotone functions $q(\theta)$.

math.ST

Toward Robust Neural Reconstruction from Sparse Point Sets

We consider the challenging problem of learning Signed Distance Functions (SDF) from sparse and noisy 3D point clouds. In contrast to recent methods that depend on smoothness priors, our method, rooted in a distributionally robust optimization (DRO) framework, incorporates a regularization term that leverages samples from the uncertainty regions of the model to improve the learned SDFs. Thanks to tractable dual formulations, we show that this framework enables a stable and efficient optimization of SDFs in the absence of ground truth supervision. Using a variety of synthetic and real data evaluations from different modalities, we show that our DRO based learning framework can improve SDF learning with respect to baselines and the state-of-the-art methods.

cs.CV

JAWS: Just A Wild Shot for Cinematic Transfer in Neural Radiance Fields

This paper presents JAWS, an optimization-driven approach that achieves the robust transfer of visual cinematic features from a reference in-the-wild video clip to a newly generated clip. To this end, we rely on an implicit-neural-representation (INR) in a way to compute a clip that shares the same cinematic features as the reference clip. We propose a general formulation of a camera optimization problem in an INR that computes extrinsic and intrinsic camera parameters as well as timing. By leveraging the differentiability of neural representations, we can back-propagate our designed cinematic losses measured on proxy estimators through a NeRF network to the proposed cinematic parameters directly. We also introduce specific enhancements such as guidance maps to improve the overall quality and efficiency. Results display the capacity of our system to replicate well known camera sequences from movies, adapting the framing, camera parameters and timing of the generated video clip to maximize the similarity with the reference clip.

cs.CV

Predictive density estimators with integrated $L_1$ loss

This paper addresses the problem of an efficient predictive density estimation for the density $q(\|y-\theta\|^2)$ of $Y$ based on $X \sim p(\|x-\theta\|^2)$ for $y, x, \theta \in \mathbb{R}^d$. The chosen criteria are integrated $L_1$ loss given by $L(\theta, \hat{q}) \, =\, \int_{\mathbb{R}^d} \big|\hat{q}(y)- q(\|y-\theta\|^2) \big| \, dy$, and the associated frequentist risk, for $\theta \in \Theta$. For absolutely continuous and strictly decreasing $q$, we establish the inevitability of scale expansion improvements $\hat{q}_c(y;X)\,=\, \frac{1}{c^d} q\big(\|y-X\|^2/c^2 \big) $ over the plug-in density $\hat{q}_1$, for a subset of values $c \in (1,c_0)$. The finding is universal with respect to $p,q$, and $d \geq 2$, and extended to loss functions $\gamma \big(L(\theta, \hat{q} ) \big)$ with strictly increasing $\gamma$. The finding is also extended to include scale expansion improvements of more general plug-in densities $q(\|y-\hat{\theta}(X)\|^2 \big)$, when the parameter space $\Theta$ is a compact subset of $\mathbb{R}^d$. Numerical analyses illustrative of the dominance findings are presented and commented upon. As a complement, we demonstrate that the unimodal assumption on $q$ is necessary with a detailed analysis of cases where the distribution of $Y|\theta$ is uniformly distributed on a ball centered about $\theta$. In such cases, we provide a univariate ($d=1$) example where the best equivariant estimator is a plug-in estimator, and we obtain cases (for $d=1,3$) where the plug-in density $\hat{q}_1$ is optimal among all $\hat{q}_c$.

math.ST

TwistSLAM++: Fusing multiple modalities for accurate dynamic semantic SLAM

Most classical SLAM systems rely on the static scene assumption, which limits their applicability in real world scenarios. Recent SLAM frameworks have been proposed to simultaneously track the camera and moving objects. However they are often unable to estimate the canonical pose of the objects and exhibit a low object tracking accuracy. To solve this problem we propose TwistSLAM++, a semantic, dynamic, SLAM system that fuses stereo images and LiDAR information. Using semantic information, we track potentially moving objects and associate them to 3D object detections in LiDAR scans to obtain their pose and size. Then, we perform registration on consecutive object scans to refine object pose estimation. Finally, object scans are used to estimate the shape of the object and constrain map points to lie on the estimated surface within the BA. We show on classical benchmarks that this fusion approach based on multimodal information improves the accuracy of object tracking.

cs.CV

TwistSLAM: Constrained SLAM in Dynamic Environment

Classical visual simultaneous localization and mapping (SLAM) algorithms usually assume the environment to be rigid. This assumption limits the applicability of those algorithms as they are unable to accurately estimate the camera poses and world structure in real life scenes containing moving objects (e.g. cars, bikes, pedestrians, etc.). To tackle this issue, we propose TwistSLAM: a semantic, dynamic and stereo SLAM system that can track dynamic objects in the environment. Our algorithm creates clusters of points according to their semantic class. Thanks to the definition of inter-cluster constraints modeled by mechanical joints (function of the semantic class), a novel constrained bundle adjustment is then able to jointly estimate both poses and velocities of moving objects along with the classical world structure and camera trajectory. We evaluate our approach on several sequences from the public KITTI dataset and demonstrate quantitatively that it improves camera and object tracking compared to state-of-the-art approaches.

cs.RO

Bayesian inference and prediction for mean-mixtures of normal distributions

We study frequentist risk properties of predictive density estimators for mean mixtures of multivariate normal distributions, involving an unknown location parameter $\theta \in \mathbb{R}^d$, and which include multivariate skew normal distributions. We provide explicit representations for Bayesian posterior and predictive densities, including the benchmark minimum risk equivariant (MRE) density, which is minimax and generalized Bayes with respect to an improper uniform density for $\theta$. For four dimensions or more, we obtain Bayesian densities that improve uniformly on the MRE density under Kullback-Leibler loss. We also provide plug-in type improvements, investigate implications for certain type of parametric restrictions on $\theta$, and illustrate and comment the findings based on numerical evaluations.

math.ST

S3LAM: Structured Scene SLAM

We propose a new SLAM system that uses the semantic segmentation of objects and structures in the scene. Semantic information is relevant as it contains high level information which may make SLAM more accurate and robust. Our contribution is twofold: i) A new SLAM system based on ORB-SLAM2 that creates a semantic map made of clusters of points corresponding to objects instances and structures in the scene. ii) A modification of the classical Bundle Adjustment formulation to constrain each cluster using geometrical priors, which improves both camera localization and reconstruction and enables a better understanding of the scene. We evaluate our approach on sequences from several public datasets and show that it improves camera pose estimation with respect to state of the art.

cs.RO

Tracking Pedestrian Heads in Dense Crowd

Tracking humans in crowded video sequences is an important constituent of visual scene understanding. Increasing crowd density challenges visibility of humans, limiting the scalability of existing pedestrian trackers to higher crowd densities. For that reason, we propose to revitalize head tracking with Crowd of Heads Dataset (CroHD), consisting of 9 sequences of 11,463 frames with over 2,276,838 heads and 5,230 tracks annotated in diverse scenes. For evaluation, we proposed a new metric, IDEucl, to measure an algorithm's efficacy in preserving a unique identity for the longest stretch in image coordinate space, thus building a correspondence between pedestrian crowd motion and the performance of a tracking algorithm. Moreover, we also propose a new head detector, HeadHunter, which is designed for small head detection in crowded scenes. We extend HeadHunter with a Particle Filter and a color histogram based re-identification module for head tracking. To establish this as a strong baseline, we compare our tracker with existing state-of-the-art pedestrian trackers on CroHD and demonstrate superiority, especially in identity preserving tracking metrics. With a light-weight head detector and a tracker which is efficient at identity preservation, we believe our contributions will serve useful in advancement of pedestrian tracking in dense crowds.

cs.CV

L6DNet: Light 6 DoF Network for Robust and Precise Object Pose Estimation with Small Datasets

Estimating the 3D pose of an object is a challenging task that can be considered within augmented reality or robotic applications. In this paper, we propose a novel approach to perform 6 DoF object pose estimation from a single RGB-D image. We adopt a hybrid pipeline in two stages: data-driven and geometric respectively. The data-driven step consists of a classification CNN to estimate the object 2D location in the image from local patches, followed by a regression CNN trained to predict the 3D location of a set of keypoints in the camera coordinate system. To extract the pose information, the geometric step consists in aligning the 3D points in the camera coordinate system with the corresponding 3D points in world coordinate system by minimizing a registration error, thus computing the pose. Our experiments on the standard dataset LineMod show that our approach is more robust and accurate than state-of-the-art methods. The approach is also validated to achieve a 6 DoF positioning task by visual servoing.

cs.CV

Visual Servoing from Deep Neural Networks

We present a deep neural network-based method to perform high-precision, robust and real-time 6 DOF visual servoing. The paper describes how to create a dataset simulating various perturbations (occlusions and lighting conditions) from a single real-world image of the scene. A convolutional neural network is fine-tuned using this dataset to estimate the relative pose between two images of the same scene. The output of the network is then employed in a visual servoing control scheme. The method converges robustly even in difficult real-world settings with strong lighting variations and occlusions.A positioning error of less than one millimeter is obtained in experiments with a 6 DOF robot.

cs.RO

On continuous distribution functions, minimax and best invariant estimators, and integrated balanced loss functions

We consider the problem of estimating a continuous distribution function $F$, as well as meaningful functions $τ(F)$ under a large class of loss functions. We obtain best invariant estimators and establish their minimaxity for Hölder continuous $τ$'s and strict bowl-shaped losses with a bounded derivative. We also introduce and motivate the use of integrated balanced loss functions which combine the criteria of an integrated distance between a decision $d$ and $F$, with the proximity of $d$ with a target estimator $d_0$. Moreover, we show how the risk analysis of procedures under such an integrated balanced loss relates to a dual risk analysis under an "unbalanced" loss, and we derive best invariant estimators, minimax estimators, risk comparisons, dominance and inadmissibility results. Finally, we expand on various illustrations and applications relative to maxima-nomination sampling, median-nomination sampling, and a case study related to bilirubin levels in the blood of babies suffering from jaundice.

math.ST

On Bayesian credible sets in restricted parameter space problems and lower bounds for frequentist coverage

For estimating a lower bounded parametric function in the framework of Marchand and Strawderman (2006), we provide through a unified approach a class of Bayesian confidence intervals with credibility $1-α$ and frequentist coverage probability bounded below by $\frac{1-α}{1+α}$. In cases where the underlying pivotal distribution is symmetric, the findings represent extensions with respect to the specification of the credible set achieved through the choice of a {\it spending function}, and include Marchand and Strawderman's HPD procedure result. For non-symmetric cases, the determination of a such a class of Bayesian credible sets fills a gap in the literature and includes an "equal-tails" modification of the HPD procedure. Several examples are presented demonstrating wide applicability.

math.ST

Estimation of a nonnegative location parameter with unknown scale

For normal canonical models, and more generally a vast array of general spherically symmetric location-scale models with a residual vector, we consider estimating the (univariate) location parameter when it is lower bounded. We provide conditions for estimators to dominate the benchmark minimax MRE estimator, and thus be minimax under scale invariant loss. These minimax estimators include the generalized Bayes estimator with respect to the truncation of the common non-informative prior onto the restricted parameter space for normal models under general convex symmetric loss, as well as non-normal models under scale invariant $L^p$ loss with $p>0$. We cover many other situations when the loss is asymmetric, and where other generalized Bayes estimators, obtained with different powers of the scale parameter in the prior measure, are proven to be minimax. We rely on various novel representations, sharp sign change analyses, as well as capitalize on Kubokawa's integral expression for risk difference technique. Several other analytical properties are obtained, including a robustness property of the generalized Bayes estimators above when the loss is either scale invariant $L^p$ or asymmetrized versions. Applications include inference in two-sample normal model with order constraints on the means.

math.ST