arXiv ScienceSearch

arXiv subjects

Aakash Kumar

Publications and source records attributed to Aakash Kumar.

At least 19 recordsLinked to original sources

Tetrahedral linkage as an intrinsic measure of glycan antifreeze behavior

Antifreeze materials prevent ice-formation by disrupting the ice-formation by binding to certain ice-planes. Cellulose, the most abundant biopolymer, has shown the ability to bind to ice-planes but the exact mechanism of this binding is far from being understood. Molecular dynamics simulations are used to investigate the hydration water of chains of cellulose-type glycans and its significance in the expression of the antifreeze behavior of sugar-derivatives found in some antifreeze materials. We find that glycans are able to prevent water from freezing near its surface by preventing their rearrangement to achieve a highly tetrahedral structure at temperatures well-below the freezing point of water. This validates our hypothesis on the role of tetrahedral coordination based on previous $\textit{ab initio}$ calculations that demonstrated cellulose prefers to bind to ice basal and prismatic planes using a tetrahedral geometry. Our findings suggest that the tetrahedral ordering of water around glycans is the key to understanding and designing cellulose-based antifreeze materials.

cond-mat.soft

A Unified Framework for Quantized and Continuous Strong Lottery Tickets

The Strong Lottery Ticket Hypothesis (SLTH) asserts that sufficiently overparameterized, randomly initialized neural networks contain sparse subnetworks that, even without any training, can match the performance of a small trained network on a given dataset. A key mathematical tool in the theoretical study of SLTH has been the Random Subset Sum Problem (RSSP). The SLTH has recently been extended to the quantized setting, where the network weights are sampled from a discrete set rather than from a continuous interval. These new results are however far from those in arbitrary-precision setting in several ways. In this work, we provide an analysis of the RSSP in the discrete setting, and use it to derive tight SLTH guarantees in the quantized case. Our analysis obtain tight bounds on the failure probability of finding a strong lottery ticket in the quantized regime, providing an exponential improvement over previous results. Most importantly, it unifies the literature by showing that both approximate representations in the continuous setting and exact representations in quantized settings naturally emerge as limiting cases of our results. This perspective not only sharpens existing bounds but also provides a cohesive framework that simultaneously handles approximation and rounding errors.

cs.LG

$\texttt{SMaSH}$ : Simplify Massive Spinor Helicity

We present $\texttt{SMaSH}$, a $\texttt{Mathematica}$ package to do spinor helicity computations in four spacetime dimensions $\href{https://github.com/aakash-kmr/SMaSH}{\text{(github)}}$. It can handle massive spinor helicity computations with explicit little group indices which is a novel feature. It can also handle massless as well as off-shell spinor helicity variables. It is designed to compute perturbative computations; it comes with predefined three point amplitudes and propagators for any masses and spins (arXiv:1709.04891). It can implement the high energy limit over an expression, check the discrete $\tt{C,P,T}$ transformations, compute contact terms and impose gauge invariance for any scattering process. We have shown the usage of such functions for computing gauge invariant Weinberg minimal amplitudes (arXiv:2506:12431, arXiv:2504:06343). The package can also generate both real and complex numerical kinematics for any $n$-point scattering for arbitrary masses and energy scales by implementing the $\tt{RAMBO}$ algorithm. It is also rich with basic spinor helicity manipulations like Schouten simplification, Clifford algebra manipulation, conversion between spinor helicity and Lorentz vectors, derivative w.r.t. spinors and their scalars, helicity scaling etc.

hep-th

Compositional Generalization in Autoregressive Models via Logit Composition

Composing autoregressive models remains a core challenge in understanding how large language models can combine behaviors or skills learned across tasks. We introduce a new and principled composition strategy for autoregressive systems, inspired by composition methods developed for diffusion models. Under a factorized-conditionals assumption, we show that the resulting composition is projective: each component model preserves control over its own designated subspace of the output distribution avoiding interference between models. This property is further preserved under smooth reparameterizations of the output space, yielding a feature-space theorem. Finally, we show that composition preserves length-generalizing behavior when the factorization assumptions and component guarantees hold uniformly at the target length. These results provide a principled understanding of when model composition and merging succeed in autoregressive systems and identify conditions under which their interactions remain stable.

cs.LG

Beyond Fixed Points: Superpolynomial Capacity of Asymmetric Hopfield Networks

Classical Hopfield networks are limited to static patterns due to symmetric weights, whereas asymmetric networks can encode temporal sequences via limit-cycle attractors. Achieving high-capacity storage of long sequences in classical synchronous asymmetric networks, however, has remained a challenge. We present a simple and robust construction within the classical asymmetric Hopfield model with binary neurons and synchronous updates, that allows $n$ neurons to support $\exp\!\big(\Omega(n/(\log n)^2)\big)$ distinct limit-cycle attractors, each with period $\exp\!\big(\Omega(\sqrt n/\log n)\big)$ and robust to random noise with flip probability up to $\frac12-o(1)$, yielding superpolynomial capacity in both the number and length of stored sequences. This is the first demonstration of such capacity for asymmetric Hopfield networks, which we obtain by combining results from combinatorics, number theory and the analysis of opinion dynamics. Our findings show that synchronous asymmetric Hopfield networks possess a sequence-memory capacity which is larger and more robust than previously recognized, demonstrating that, in both biological and artificial neural systems, robust sequence representation can be achieved through coarse architectural motifs rather than complex nonlinearities.

cs.LG

Quantization vs Pruning: Insights from the Strong Lottery Ticket Hypothesis

Quantization is an essential technique for making neural networks more efficient, yet our theoretical understanding of it remains limited. Previous works demonstrated that extremely low-precision networks, such as binary networks, can be constructed by pruning large, randomly-initialized networks, and showed that the ratio between the size of the original and the pruned networks is at most polylogarithmic. The specific pruning method they employed inspired a line of theoretical work known as the Strong Lottery Ticket Hypothesis (SLTH), which leverages insights from the Random Subset Sum Problem. However, these results primarily address the continuous setting and cannot be applied to extend SLTH results to the quantized setting. In this work, we build on foundational results by Borgs et al. on the Number Partitioning Problem to derive new theoretical results for the Random Subset Sum Problem in a quantized setting. Using these results, we then extend the SLTH framework to finite-precision networks. While prior work on SLTH showed that pruning allows approximation of a certain class of neural networks, we demonstrate that, in the quantized setting, the analogous class of target discrete neural networks can be represented exactly, and we prove optimal bounds on the necessary overparameterization of the initial network as a function of the precision of the target network.

cs.LG

Compton amplitude and Contact term(s) in the Spinor Helicity formalism

In gauge theories, contact terms play an important role in ensuring gauge invariance. In the spinor helicity formalism, the choice of a gauge-fixing condition manifests itself in the form of the choice of reference vector to write the massless polarization vector(s). However, this choice must be irrelevant in any gauge-invariant observable. We use this principle to determine contact term for Electromagnetic Compton amplitude. We considered three-point function between two massive particles \& a photon to be one which is responsible for soft photon theorem/Coulomb and demonstrate that it is possible to use the above-mentioned principle to find the contact term for the tree-level Compton amplitude of two bosonic massive spinning particles and two photons. The final result does not suffer from any spurious poles.

hep-th

Compton amplitude for massive bosons of arbitrary spin

In this work, we write down an analytic expression of electromagnetic tree-level Compton amplitude for a completely symmetric traceless (bosonic) higher spin particle in any dimension. Our analysis is restricted to the three-point function, which is unique in the Infrared and responsible for the Coulomb interactions/soft photon theorem. We propose an analogue of $R_\xi$ gauge in a theory of higher spin particles. We demonstrate that the theory is unitary only for $\xi=1,\infty$.

hep-th

Revealing the impact of synthetic native samples and multi-tasking strategies in Hindi-English code-mixed humour and sarcasm detection

In this paper, we reported our experiments with various strategies to improve code-mixed humour and sarcasm detection. Particularly, we tried three approaches: (i) native sample mixing, (ii) multi-task learning (MTL), and (iii) prompting and instruction finetuning very large multilingual language models (VMLMs). In native sample mixing, we added monolingual task samples to code-mixed training sets. In MTL learning, we relied on native and code-mixed samples of a semantically related task (hate detection in our case). Finally, in our third approach, we evaluated the efficacy of VMLMs via few-shot context prompting and instruction finetuning. Some interesting findings we got are (i) adding native samples improved humor (raising the F1-score up to 6.76%) and sarcasm (raising the F1-score up to 8.64%) detection, (ii) training MLMs in an MTL framework boosted performance for both humour (raising the F1-score up to 10.67%) and sarcasm (increment up to 12.35% in F1-score) detection, and (iii) prompting and instruction finetuning VMLMs couldn't outperform the other approaches. Finally, our ablation studies and error analysis discovered the cases where our model is yet to improve. We provided our code for reproducibility.

cs.CL

Improving code-mixed hate detection by native sample mixing: A case study for Hindi-English code-mixed scenario

Hate detection has long been a challenging task for the NLP community. The task becomes complex in a code-mixed environment because the models must understand the context and the hate expressed through language alteration. Compared to the monolingual setup, we see much less work on code-mixed hate as large-scale annotated hate corpora are unavailable for the study. To overcome this bottleneck, we propose using native language hate samples (native language samples/ native samples hereafter). We hypothesise that in the era of multilingual language models (MLMs), hate in code-mixed settings can be detected by majorly relying on the native language samples. Even though the NLP literature reports the effectiveness of MLMs on hate detection in many cross-lingual settings, their extensive evaluation in a code-mixed scenario is yet to be done. This paper attempts to fill this gap through rigorous empirical experiments. We considered the Hindi-English code-mixed setup as a case study as we have the linguistic expertise for the same. Some of the interesting observations we got are: (i) adding native hate samples in the code-mixed training set, even in small quantity, improved the performance of MLMs for code-mixed hate detection, (ii) MLMs trained with native samples alone observed to be detecting code-mixed hate to a large extent, (iii) the visualisation of attention scores revealed that, when native samples were included in training, MLMs could better focus on the hate emitting words in the code-mixed context, and (iv) finally, when hate is subjective or sarcastic, naively mixing native samples doesn't help much to detect code-mixed hate. We will release the data and code repository to reproduce the reported results.

cs.CL

Sparse Points to Dense Clouds: Enhancing 3D Detection with Limited LiDAR Data

3D detection is a critical task that enables machines to identify and locate objects in three-dimensional space. It has a broad range of applications in several fields, including autonomous driving, robotics and augmented reality. Monocular 3D detection is attractive as it requires only a single camera, however, it lacks the accuracy and robustness required for real world applications. High resolution LiDAR on the other hand, can be expensive and lead to interference problems in heavy traffic given their active transmissions. We propose a balanced approach that combines the advantages of monocular and point cloud-based 3D detection. Our method requires only a small number of 3D points, that can be obtained from a low-cost, low-resolution sensor. Specifically, we use only 512 points, which is just 1% of a full LiDAR frame in the KITTI dataset. Our method reconstructs a complete 3D point cloud from this limited 3D information combined with a single image. The reconstructed 3D point cloud and corresponding image can be used by any multi-modal off-the-shelf detector for 3D object detection. By using the proposed network architecture with an off-the-shelf multi-modal 3D detector, the accuracy of 3D detection improves by 20% compared to the state-of-the-art monocular detection methods and 6% to 9% compare to the baseline multi-modal methods on KITTI and JackRabbot datasets.

cs.CV

Three point interaction of Dirac fermions with higher spin particles and discrete symmetries

We constructed all possible kinematically allowed three-point interactions of two massless Dirac spinors with massive higher-spin bosons. In any $D$ spacetime, the interactions have been constructed using the projections of the higher spin irreducible representations of $Spin(D-1)$ over the product of two irreducible spinor representations of $Spin(D-2)$. Based on this analysis, we have further classified the space of theories involving two massless Dirac spinors and a single (or multiple) massive higher spin(s) based on the discrete symmetries: $C,\, R,$ and $ T$. We found that in any $D=2m+1/2m$, the interacting theories of a single massive higher spin have a \enquote{$m$} mod $2$ (or $D$ mod $4$) classification.

hep-th

Advanced Efficient Strategy for Detection of Dark Objects Based on Spiking Network with Multi-Box Detection

Several deep learning algorithms have shown amazing performance for existing object detection tasks, but recognizing darker objects is the largest challenge. Moreover, those techniques struggled to detect or had a slow recognition rate, resulting in significant performance losses. As a result, an improved and accurate detection approach is required to address the above difficulty. The whole study proposes a combination of spiked and normal convolution layers as an energy-efficient and reliable object detector model. The proposed model is split into two sections. The first section is developed as a feature extractor, which utilizes pre-trained VGG16, and the second section of the proposal structure is the combination of spiked and normal Convolutional layers to detect the bounding boxes of images. We drew a pre-trained model for classifying detected objects. With state of the art Python libraries, spike layers can be trained efficiently. The proposed spike convolutional object detector (SCOD) has been evaluated on VOC and Ex-Dark datasets. SCOD reached 66.01% and 41.25% mAP for detecting 20 different objects in the VOC-12 and 12 objects in the Ex-Dark dataset. SCOD uses 14 Giga FLOPS for its forward path calculations. Experimental results indicated superior performance compared to Tiny YOLO, Spike YOLO, YOLO-LITE, Tinier YOLO and Center of loc+Xception based on mAP for the VOC dataset.

cs.CV

Applying adversarial networks to increase the data efficiency and reliability of Self-Driving Cars

Convolutional Neural Networks (CNNs) are vulnerable to misclassifying images when small perturbations are present. With the increasing prevalence of CNNs in self-driving cars, it is vital to ensure these algorithms are robust to prevent collisions from occurring due to failure in recognizing a situation. In the Adversarial Self-Driving framework, a Generative Adversarial Network (GAN) is implemented to generate realistic perturbations in an image that cause a classifier CNN to misclassify data. This perturbed data is then used to train the classifier CNN further. The Adversarial Self-driving framework is applied to an image classification algorithm to improve the classification accuracy on perturbed images and is later applied to train a self-driving car to drive in a simulation. A small-scale self-driving car is also built to drive around a track and classify signs. The Adversarial Self-driving framework produces perturbed images through learning a dataset, as a result removing the need to train on significant amounts of data. Experiments demonstrate that the Adversarial Self-driving framework identifies situations where CNNs are vulnerable to perturbations and generates new examples of these situations for the CNN to train on. The additional data generated by the Adversarial Self-driving framework provides sufficient data for the CNN to generalize to the environment. Therefore, it is a viable tool to increase the resilience of CNNs to perturbations. Particularly, in the real-world self-driving car, the application of the Adversarial Self-Driving framework resulted in an 18 % increase in accuracy, and the simulated self-driving model had no collisions in 30 minutes of driving.

cs.CV

A Gapped Phase in Semimetallic T$_{d}$-WTe$_{2}$ Induced by Lithium Intercalation

The Weyl semimetal WTe$_{2}$ has shown several correlated electronic behaviors, such as the quantum spin Hall effect, superconductivity, ferroelectricity, and a possible exciton insulator state, all of which can be tuned by various physical and chemical approaches. Here, we discover a new electronic phase in WTe$_{2}$ induced by lithium intercalation. The new phase exhibits an increasing resistivity with decreasing temperature and its carrier density is almost two orders of magnitude lower than the carrier density of the semi-metallic T$_{d}$ phase, probed by in situ Hall measurements as a function of lithium intercalation. Our theoretical calculations predict the new lithiated phase to be a charge density wave (CDW) phase with a bandgap of ~ 0.14 eV, in good agreement with the in situ transport data. The new phase is structurally distinct from the initial T$_{d}$ phase, characterized by polarization angle-dependent Raman spectroscopy, and large lattice distortions close to 6 % are predicted in the new phase. Thus, we report the first experimental evidence of CDW in T$_{d}$-WTe$_{2}$, projecting WTe$_{2}$ as a new playground for studying the interplay between CDW and superconductivity. Our finding of a new gapped phase in a two-dimensional (2D) semi-metal also demonstrates electrochemical intercalation as a powerful tuning knob for modulating electron density and phase stability in 2D materials.

cond-mat.mtrl-sci

Heterointerface control over lithium-induced phase transitions in MoS2 heterostructures

Phase transitions of two-dimensional materials and their heterostructures enable many applications including electrochemical energy storage, catalysis, and memory; however, the nucleation pathways by which these transitions proceed remain underexplored, prohibiting engineering control for these applications. Here, we demonstrate that the lithium intercalation-induced 2H-1T' phase transition in MoS2 proceeds via nucleation of the 1T' phase at a heterointerface by monitoring the phase transition of MoS2/graphene and MoS2/hexagonal boron nitride (hBN) heterostructures with Raman spectroscopy in situ during intercalation. We observe that graphene-MoS2 heterointerfaces require an increase of 0.8 V in applied electrochemical potential to nucleate the 1T' phase in MoS2 compared to hBN-MoS2 heterointerfaces. The increased nucleation barrier at graphene-MoS2 heterointerfaces is due to the reduced charge transfer from lithium to MoS2 at the heterointerface as lithium also dopes graphene based on ab initio calculations. Further, we show that the growth of the 1T' domain propagates along the heterointerface, rather than through the interior of MoS2. Our results provide the first experimental observations of the heterogeneous nucleation and growth of intercalation-induced phase transitions in two-dimensional materials and heterointerface effects on their phase transition.

cond-mat.mtrl-sci

PC-DAN: Point Cloud based Deep Affinity Network for 3D Multi-Object Tracking (Accepted as an extended abstract in JRDB-ACT Workshop at CVPR21)

In recent times, the scope of LIDAR (Light Detection and Ranging) sensor-based technology has spread across numerous fields. It is popularly used to map terrain and navigation information into reliable 3D point cloud data, potentially revolutionizing the autonomous vehicles and assistive robotic industry. A point cloud is a dense compilation of spatial data in 3D coordinates. It plays a vital role in modeling complex real-world scenes since it preserves structural information and avoids perspective distortion, unlike image data, which is the projection of a 3D structure on a 2D plane. In order to leverage the intrinsic capabilities of the LIDAR data, we propose a PointNet-based approach for 3D Multi-Object Tracking (MOT).

cs.CV

Heterointerface effects on lithium-induced phase transitions in intercalated MoS2

The intercalation-induced phase transition of MoS2 from the semiconducting 2H to the semimetallic 1T' phase has been studied in detail for nearly a decade; however, the effects of a heterointerface between MoS2 and other two-dimensional (2D) crystals on the phase transition have largely been overlooked. Here, ab initio calculations show that intercalating Li at a MoS2-hexagonal boron nitride (hBN) interface stabilizes the 1T phase over the 2H phase of MoS2 by ~ 100 mJ m-2, suggesting that encapsulating MoS2 with hBN may lower the electrochemical energy needed for the intercalation-induced phase transition. However, in situ Raman spectroscopy of hBN-MoS2-hBN heterostructures during electrochemical intercalation of Li+ shows that the phase transition occurs at the same applied voltage for the heterostructure as for bare MoS2. We hypothesize that the predicted thermodynamic stabilization of the 1T'-MoS2-hBN interface is counteracted by an energy barrier to the phase transition imposed by the steric hindrance of the heterointerface. The phase transition occurs at lower applied voltages upon heating the heterostructure, which supports our hypothesis. Our study highlights that interfacial effects of 2D heterostructures can go beyond modulating electrical properties and can modify electrochemical and phase transition behaviors.

cond-mat.mtrl-sci