arXiv ScienceSearch

arXiv subjects

Xi Qu

Publications and source records attributed to Xi Qu.

5 recordsLinked to original sources

Semi-nonparametric estimation of spatial dynamic panel data models with nonparametric spatial weights

We develop a semi-nonparametric framework for spatial dynamic panel data (SDPD) models with two-way fixed effects when the spatial interaction structure is unknown beyond a distance measure. This is accomplished by modelling spatial weights in the outcome, lagged-outcome, and disturbance channels as unknown functions of underlying economic distances. These enter the SDPD system through matrix-function operators, providing a unified approach that accommodates both spatial autoregressive and matrix exponential spatial specifications. Allowing for unknown heteroskedasticity, we propose sieve GMM estimators based on a stacked set of linear and quadratic moment conditions, and derive a feasible optimal GMM estimator and a more efficient feasible best GMM estimator. As $(n, T) \rightarrow \infty$, the parametric component is $\sqrt{n(T - 1)}$-consistent and asymptotically normal, echoing classical semi-nonparametric results. Monte Carlo experiments indicate excellent finite-sample performance. We apply the method to 'witch' killings as studied by Miguel (2005), and find that economic-geography proximity rather than cultural-geography proximity between communities significantly amplifies spatial dependence in these economic murders.

econ.EM

Estimating Social Network Models with Link Misclassification

We propose an adjusted 2SLS estimator for social network models when reported binary network links are misclassified (some zeros reported as ones and vice versa) due, e.g., to survey respondents' recall errors, or lapses in data input. We show misclassification adds new sources of correlation between the regressors and errors, which makes all covariates endogenous and invalidates conventional estimators. We resolve these issues by constructing a novel estimator of misclassification rates and using those estimates to both adjust endogenous peer outcomes and construct new instruments for 2SLS estimation. A distinctive feature of our method is that it does not require structural modeling of link formation. Simulation results confirm our adjusted 2SLS estimator corrects the bias from a naive, unadjusted 2SLS estimator which ignores misclassification and uses conventional instruments. We apply our method to study peer effects in household decisions to participate in a microfinance program in Indian villages.

econ.EM

Wald inference on varying coefficients

We present simple to implement Wald-type statistics that deliver a general nonparametric inference theory for linear restrictions on varying coefficients in a range of regression models allowing for cross-sectional or spatial dependence. We provide a general central limit theorem that covers a broad range of error spatial dependence structures, allows for a degree of misspecification robustness via nonparametric spatial weights and permits inference on both varying regression and spatial dependence parameters. Using our method, we first uncover evidence of constant returns to scale in the Chinese nonmetal mineral industry's production function, and then show that Boston house prices respond nonlinearly to proximity to employment centers. A simulation study confirms that our tests perform very well in finite samples.

econ.EM

Consistent specification testing under spatial dependence

We propose a series-based nonparametric specification test for a regression function when data are spatially dependent, the `space' being of a general economic or social nature. Dependence can be parametric, parametric with increasing dimension, semiparametric or any combination thereof, thus covering a vast variety of settings. These include spatial error models of varying types and levels of complexity. Under a new smooth spatial dependence condition, our test statistic is asymptotically standard normal. To prove the latter property, we establish a central limit theorem for quadratic forms in linear processes in an increasing dimension setting. Finite sample performance is investigated in a simulation study, with a bootstrap method also justified and illustrated, and empirical examples illustrate the test with real-world data.

econ.EM

Clust-LDA: Joint Model for Text Mining and Author Group Inference

Social media corpora pose unique challenges and opportunities, including typically short document lengths and rich meta-data such as author characteristics and relationships. This creates great potential for systematic analysis of the enormous body of the users and thus provides implications for industrial strategies such as targeted marketing. Here we propose a novel and statistically principled method, clust-LDA, which incorporates authorship structure into the topical modeling, thus accomplishing the task of the topical inferences across documents on the basis of authorship and, simultaneously, the identification of groupings between authors. We develop an inference procedure for clust-LDA and demonstrate its performance on simulated data, showing that clust-LDA out-performs the "vanilla" LDA on the topic identification task where authors exhibit distinctive topical preference. We also showcase the empirical performance of clust-LDA based on a real-world social media dataset from Reddit.

cs.IR