arXiv ScienceSearch

arXiv subjects

Linlin Liu

Publications and source records attributed to Linlin Liu.

At least 19 recordsLinked to original sources

Reversible fully spin polarization in strain-engineered two-dimensional fully compensated magnets

Achieving controllable spin polarization and its reversal in symmetry-compensated magnets. Here we demonstrate, using symmetry analysis and a minimal tight-binding model, that uniaxial strain removes these constraints by inducing inequivalence between magnetic sublattices in two-dimensional (2D) system, driving an altermagnetic (AM) state into a fully compensated ferrimagnetic (fFIM) state and enabling fully spin polarization. Furthermore, strain along orthogonal directions gives rise to two energetically degenerate fFIM states with opposite spin polarization, enabling reversible spin switching. More importantly, the two symmetry-related fFIM states can be regarded as distinct ferroelastic variants, suggesting that this model or mechanism can be extended to ferroelastic fFIM systems. The generality of this mechanism is confirmed by combining spin-group analysis, first-principles calculations, and Boltzmann transport theory in representative candidates, including AM Mn$_2$SeO and ferroelastic fFIM V$_2$SO. Our results reveal a universal symmetry-driven framework for strain-controlled and -reversible fully spin-polarized transport and identify strain-engineered AM and ferroelastic fFIM systems as a promising platform for volatile and nonvolatile spintronic applications.

cond-mat.mtrl-sci

Associative $H$-pseudoalgebras with a semigroup

Family algebraic structures indexed by a semigroup arise naturally in renormalizations of quantum field theory. In this paper, we first define the notion of $\Omega$-associative $H$-pseudoalgebra, where the operations are indexed by pairs of elements from a semigroup $\Omega$. Then we construct $\Omega$-associative $H$-pseudoalgebras from associative $H$-pseudoalgebras, $\Omega$-associative algebras, Rota-Baxter family algebras, $\Omega$-type $H$-pseudoalgebras and family-type $H$-pseudoalgebras. Moreover, we investigate the cohomology of $\Omega$-associative $H$-pseudoalgebras and establish that it both induces the cohomology of pseudo-$\mathcal{O}$-operator families and governs the associated formal deformations. As an application, we show that the first-order deformation of a commutative $\Omega$-associative $H$-pseudoalgebra yields an $\Omega$-Poisson $H$-pseudoalgebra.

math.RA

Research on the Crystal Growth, Band Structure and Luminescence Mechanism of (CH3NH3)2HgI4

Nuclear radiation detectors play a crucial role in fields such as nuclear safety and medical imaging. The core of their performance lies in the selection of detection materials. Semiconductor detectors have become a hot topic in current research due to their advantages such as small size, good energy resolution, and high detection efficiency. As one of the most promising materials for fabricating room - temperature nuclear radiation semiconductor detectors, HgI2 exhibits excellent detection performance due to its high atomic number, large band gap, strong ray - stopping power, and high volume dark resistivity. However, issues such as poor chemical stability and low vacancy mobility of HgI2 limit its development. Therefore, researchers have carried out inorganic doping/organic hybridization on it. By introducing the organic ligand CH3NH3I, the synthesis of organic - inorganic hybrid compounds based on HgI2 is expected to significantly improve the stability of HgI2. Research on organic - inorganic hybrid metal halide crystals shows that this material has great application potential in the field of luminescent materials.

cond-mat.mtrl-sci

Two-Dimensional Graphene-like BeO Sheet: A Promising Deep-Ultraviolet Nonlinear Optical Materials System with Strong and Highly Tunable Second Harmonic Generation

Two-dimensional (2D) materials with large band gaps and strong and tunable second-harmonic generation (SHG) coefficients play an important role in the miniaturization of deep-ultraviolet (DUV) nonlinear optical (NLO) devices. Despite the existence of numerous experimentally synthesized 2D materials, none of them have been reported to meet DUV NLO requirements. Herein, to the first time, an experimentally available graphene-like BeO monolayer only formed by NLO-active [BeO3] unit is suggested as a promising 2D DUV NLO material due to its ultrawide band gap (6.86 eV) and a strong SHG effect (\{chi}_"22" ^((2))(2D) = 6.81 {\AA}\times pm/V) based on the first-principles calculations. By applying stacking, strain, and twist engineering methods, several 2D BeO sheets have been predicted, and the flexible structural characteristics endow them with tunable NLO properties. Remarkably, the extremely stress-sensitive out-of-plane \{chi}_"15" ^((2))(2D) and \{chi}_"33" ^((2))(2D) (exceptional 30% change) and the robust in-plane \{chi}_"22" ^((2))(2D) against large strains can be achieved together in AC-, AAC-, AAE, and ACE-stacking BeO sheets under in-plane biaxial strain, exhibiting emergent phenomena uniquely not yet seen in other known 2D NLO materials. Our present results reveal that 2D BeO systems should be a new option for 2D DUV NLO materials.

cond-mat.mes-hall

Learning to Synthesize Compatible Fashion Items Using Semantic Alignment and Collocation Classification: An Outfit Generation Framework

The field of fashion compatibility learning has attracted great attention from both the academic and industrial communities in recent years. Many studies have been carried out for fashion compatibility prediction, collocated outfit recommendation, artificial intelligence (AI)-enabled compatible fashion design, and related topics. In particular, AI-enabled compatible fashion design can be used to synthesize compatible fashion items or outfits in order to improve the design experience for designers or the efficacy of recommendations for customers. However, previous generative models for collocated fashion synthesis have generally focused on the image-to-image translation between fashion items of upper and lower clothing. In this paper, we propose a novel outfit generation framework, i.e., OutfitGAN, with the aim of synthesizing a set of complementary items to compose an entire outfit, given one extant fashion item and reference masks of target synthesized items. OutfitGAN includes a semantic alignment module, which is responsible for characterizing the mapping correspondence between the existing fashion items and the synthesized ones, to improve the quality of the synthesized images, and a collocation classification module, which is used to improve the compatibility of a synthesized outfit. In order to evaluate the performance of our proposed models, we built a large-scale dataset consisting of 20,000 fashion outfits. Extensive experimental results on this dataset show that our OutfitGAN can synthesize photo-realistic outfits and outperform state-of-the-art methods in terms of similarity, authenticity and compatibility measurements.

cs.LG

Multi-view X-ray Image Synthesis with Multiple Domain Disentanglement from CT Scans

X-ray images play a vital role in the intraoperative processes due to their high resolution and fast imaging speed and greatly promote the subsequent segmentation, registration and reconstruction. However, over-dosed X-rays superimpose potential risks to human health to some extent. Data-driven algorithms from volume scans to X-ray images are restricted by the scarcity of paired X-ray and volume data. Existing methods are mainly realized by modelling the whole X-ray imaging procedure. In this study, we propose a learning-based approach termed CT2X-GAN to synthesize the X-ray images in an end-to-end manner using the content and style disentanglement from three different image domains. Our method decouples the anatomical structure information from CT scans and style information from unpaired real X-ray images/ digital reconstructed radiography (DRR) images via a series of decoupling encoders. Additionally, we introduce a novel consistency regularization term to improve the stylistic resemblance between synthesized X-ray images and real X-ray images. Meanwhile, we also impose a supervised process by computing the similarity of computed real DRR and synthesized DRR images. We further develop a pose attention module to fully strengthen the comprehensive information in the decoupled content code from CT scans, facilitating high-quality multi-view image synthesis in the lower 2D space. Extensive experiments were conducted on the publicly available CTSpine1K dataset and achieved 97.8350, 0.0842 and 3.0938 in terms of FID, KID and defined user-scored X-ray similarity, respectively. In comparison with 3D-aware methods ($\pi$-GAN, EG3D), CT2X-GAN is superior in improving the synthesis quality and realistic to the real X-ray images.

eess.IV

Is GPT-3 a Good Data Annotator?

Data annotation is the process of labeling data that could be used to train machine learning models. Having high-quality annotation is crucial, as it allows the model to learn the relationship between the input data and the desired output. GPT-3, a large-scale language model developed by OpenAI, has demonstrated impressive zero- and few-shot performance on a wide range of NLP tasks. It is therefore natural to wonder whether it can be used to effectively annotate data for NLP tasks. In this paper, we evaluate the performance of GPT-3 as a data annotator by comparing it with traditional data annotation methods and analyzing its output on a range of tasks. Through this analysis, we aim to provide insight into the potential of GPT-3 as a general-purpose data annotator in NLP.

cs.CL

Pressured-induced superconductivity extending across the topological phase transition in thallium-based topological materials TlBi(S1-xSex)2

The coexistence of superconductivity and topology holds the potential to realize exotic quantum states of matter. Here we report that superconductivity induced by high pressure in three thallium-based materials, covering the phase transition from a normal insulator (TlBiS2) to a topological insulator (TlBiSe2) through a Dirac semimetal (TlBiSeS). By increasing the pressure up to 60 GPa, we observe superconductivity phase diagrams with maximal Tc values at 6.0-8.1 K. Our density-functional theory calculations reveal topological surface states in superconductivity phases for all three compounds. Our study paves the path to explore topological superconductivity and topological phase transitions.

cond-mat.supr-con

Towards Robust Low-Resource Fine-Tuning with Multi-View Compressed Representations

Due to the huge amount of parameters, fine-tuning of pretrained language models (PLMs) is prone to overfitting in the low resource scenarios. In this work, we present a novel method that operates on the hidden representations of a PLM to reduce overfitting. During fine-tuning, our method inserts random autoencoders between the hidden layers of a PLM, which transform activations from the previous layers into multi-view compressed representations before feeding them into the upper layers. The autoencoders are plugged out after fine-tuning, so our method does not add extra parameters or increase computation cost during inference. Our method demonstrates promising performance improvement across a wide range of sequence- and token-level low-resource NLP tasks.

cs.CL

Generic Cryo-CMOS Device Modeling and EDACompatible Platform for Reliable Cryogenic IC Design

This paper outlines the establishment of a generic cryogenic CMOS database in which key electrical parameters and transfer characteristics of the MOSFETs are quantified as functions of device size, temperature/frequency responses. Meanwhile, comprehensive device statistical study is conducted to evaluate the influence of variation and mismatch effects at low temperatures. Furthermore, by incorporating the Cryo-CMOS compact model into the process design kit (PDK), the cryogenic 4 Kb SRAM, 5-bit flash ADC and 8-bit current steering DAC are designed, and their performance is readily investigated and optimized on the EDA-compatible platform, hence laying a solid foundation for large-scale cryogenic IC design.

eess.SY

Hierarchical Vectorization for Portrait Images

Aiming at developing intuitive and easy-to-use portrait editing tools, we propose a novel vectorization method that can automatically convert raster images into a 3-tier hierarchical representation. The base layer consists of a set of sparse diffusion curves (DC) which characterize salient geometric features and low-frequency colors and provide means for semantic color transfer and facial expression editing. The middle level encodes specular highlights and shadows to large and editable Poisson regions (PR) and allows the user to directly adjust illumination via tuning the strength and/or changing shape of PR. The top level contains two types of pixel-sized PRs for high-frequency residuals and fine details such as pimples and pigmentation. We also train a deep generative model that can produce high-frequency residuals automatically. Thanks to the meaningful organization of vector primitives, editing portraits becomes easy and intuitive. In particular, our method supports color transfer, facial expression editing, highlight and shadow editing and automatic retouching. Thanks to the linearity of the Laplace operator, we introduce alpha blending, linear dodge and linear burn to vector editing and show that they are effective in editing highlights and shadows. To quantitatively evaluate the results, we extend the commonly used FLIP metric (which measures differences between two images) by considering illumination. The new metric, called illumination-sensitive FLIP or IS-FLIP, can effectively capture the salient changes in color transfer results, and is more consistent with human perception than FLIP and other quality measures on portrait images. We evaluate our method on the FFHQR dataset and show that our method is effective for common portrait editing tasks, such as retouching, light editing, color transfer and expression editing. We will make the code and trained models publicly available.

cs.CV

Flexible Portrait Image Editing with Fine-Grained Control

We develop a new method for portrait image editing, which supports fine-grained editing of geometries, colors, lights and shadows using a single neural network model. We adopt a novel asymmetric conditional GAN architecture: the generators take the transformed conditional inputs, such as edge maps, color palette, sliders and masks, that can be directly edited by the user; the discriminators take the conditional inputs in the way that can guide controllable image generation more effectively. Taking color editing as an example, we feed color palettes (which can be edited easily) into the generator, and color maps (which contain positional information of colors) into the discriminator. We also design a region-weighted discriminator so that higher weights are assigned to more important regions, like eyes and skin. Using a color palette, the user can directly specify the desired colors of hair, skin, eyes, lip and background. Color sliders allow the user to blend colors in an intuitive manner. The user can also edit lights and shadows by modifying the corresponding masks. We demonstrate the effectiveness of our method by evaluating it on the CelebAMask-HQ dataset with a wide range of tasks, including geometry/color/shadow/light editing, hand-drawn sketch to image translation, and color transfer. We also present ablation studies to justify our design.

cs.CV

Enhancing Multilingual Language Model with Massive Multilingual Knowledge Triples

Knowledge-enhanced language representation learning has shown promising results across various knowledge-intensive NLP tasks. However, prior methods are limited in efficient utilization of multilingual knowledge graph (KG) data for language model (LM) pretraining. They often train LMs with KGs in indirect ways, relying on extra entity/relation embeddings to facilitate knowledge injection. In this work, we explore methods to make better use of the multilingual annotation and language agnostic property of KG triples, and present novel knowledge based multilingual language models (KMLMs) trained directly on the knowledge triples. We first generate a large amount of multilingual synthetic sentences using the Wikidata KG triples. Then based on the intra- and inter-sentence structures of the generated data, we design pretraining tasks to enable the LMs to not only memorize the factual knowledge but also learn useful logical patterns. Our pretrained KMLMs demonstrate significant performance improvements on a wide range of knowledge-intensive cross-lingual tasks, including named entity recognition (NER), factual knowledge retrieval, relation classification, and a newly designed logical reasoning task.

cs.CL

On the Effectiveness of Adapter-based Tuning for Pretrained Language Model Adaptation

Adapter-based tuning has recently arisen as an alternative to fine-tuning. It works by adding light-weight adapter modules to a pretrained language model (PrLM) and only updating the parameters of adapter modules when learning on a downstream task. As such, it adds only a few trainable parameters per new task, allowing a high degree of parameter sharing. Prior studies have shown that adapter-based tuning often achieves comparable results to fine-tuning. However, existing work only focuses on the parameter-efficient aspect of adapter-based tuning while lacking further investigation on its effectiveness. In this paper, we study the latter. We first show that adapter-based tuning better mitigates forgetting issues than fine-tuning since it yields representations with less deviation from those generated by the initial PrLM. We then empirically compare the two tuning methods on several downstream NLP tasks and settings. We demonstrate that 1) adapter-based tuning outperforms fine-tuning on low-resource and cross-lingual tasks; 2) it is more robust to overfitting and less sensitive to changes in learning rates.

cs.CL

Towards Multi-Sense Cross-Lingual Alignment of Contextual Embeddings

Cross-lingual word embeddings (CLWE) have been proven useful in many cross-lingual tasks. However, most existing approaches to learn CLWE including the ones with contextual embeddings are sense agnostic. In this work, we propose a novel framework to align contextual embeddings at the sense level by leveraging cross-lingual signal from bilingual dictionaries only. We operationalize our framework by first proposing a novel sense-aware cross entropy loss to model word senses explicitly. The monolingual ELMo and BERT models pretrained with our sense-aware cross entropy loss demonstrate significant performance improvement for word sense disambiguation tasks. We then propose a sense alignment objective on top of the sense-aware cross entropy loss for cross-lingual model pretraining, and pretrain cross-lingual models for several language pairs (English to German/Spanish/Japanese/Chinese). Compared with the best baseline results, our cross-lingual models achieve 0.52%, 2.09% and 1.29% average performance improvements on zero-shot cross-lingual NER, sentiment classification and XNLI tasks, respectively.

cs.CL

A paper's corresponding affiliation and first affiliation are consistent at the country level in Web of Science

The purpose of this study is to explore the relationship between the first affiliation and the corresponding affiliation at the different levels via the scientometric analysis We select over 18 million papers in the core collection database of Web of Science (WoS) published from 2000 to 2015, and measure the percentage of match between the first and the corresponding affiliation at the country and institution level. We find that a paper's the first affiliation and the corresponding affiliation are highly consistent at the country level, with over 98% of the match on average. However, the match at the institution level is much lower, which varies significantly with time and country. Hence, for studies at the country level, using the first and corresponding affiliations are almost the same. But we may need to take more cautions to select affiliation when the institution is the focus of the investigation. In the meanwhile, we find some evidence that the recorded corresponding information in the WoS database has undergone some changes since 2013, which sheds light on future studies on the comparison of different databases or the affiliation accuracy of WoS. Our finding relies on the records of WoS, which may not be entirely accurate. Given the scale of the analysis, our findings can serve as a useful reference for further studies when country allocation or institute allocation is needed. Existing studies on comparisons of straight counting methods usually cover a limited number of papers, a particular research field or a limited range of time. More importantly, using the number counted can not sufficiently tell if the corresponding and first affiliation are similar. This paper uses a metric similar to Jaccard similarity to measure the percentage of the match and performs a comprehensive analysis based on a large-scale bibliometric database.

cs.DL

DAGA: Data Augmentation with a Generation Approach for Low-resource Tagging Tasks

Data augmentation techniques have been widely used to improve machine learning performance as they enhance the generalization capability of models. In this work, to generate high quality synthetic data for low-resource tagging tasks, we propose a novel augmentation method with language models trained on the linearized labeled sentences. Our method is applicable to both supervised and semi-supervised settings. For the supervised settings, we conduct extensive experiments on named entity recognition (NER), part of speech (POS) tagging and end-to-end target based sentiment analysis (E2E-TBSA) tasks. For the semi-supervised settings, we evaluate our method on the NER task under the conditions of given unlabeled data only and unlabeled data plus a knowledge base. The results show that our method can consistently outperform the baselines, particularly when the given gold training data are less.

cs.CL

An associative analogy of Lie H-pseudobialgebra

The purpose of this paper is to study infinitesimal H-pseudobialgebra, which is an associative analogy of Lie H-pseudobialgebra. We first define the infinitesimal H-pseudobialgebra and investigate some properties of this new algebraic structure. Then we consider the coboundary infinitesimal H-pseudobialgebra, which is the subclass of infinitesimal H-pseudobialgebra and we obtain the associative Yang-Baxter equation over an associative H-pseudoalgebra. Finally, we found the connection between the (coboundary) infinitesimal H-pseudobialgebra and the (coboundary) Lie H-pseudobialgebra. Meanwhile, the relationship between the associative Yang-Baxter equation and the classical Yang-Baxter equation (over an H-pseudoalgebra) is established.

math.RA