arXiv ScienceSearch

arXiv subjects

Yoonji Lee

Publications and source records attributed to Yoonji Lee.

7 recordsLinked to original sources

HyCal: A Training-Free Prototype Calibration Method for Cross-Discipline Few-Shot Class-Incremental Learning

Pretrained Vision-Language Models (VLMs) like CLIP show promise in continual learning, but existing Few-Shot Class-Incremental Learning (FSCIL) methods assume homogeneous domains and balanced data distributions, limiting real-world applicability where data arises from heterogeneous disciplines with imbalanced sample availability and varying visual complexity. We identify Domain Gravity, a representational asymmetry where data imbalance across heterogeneous domains causes overrepresented or low-entropy domains to disproportionately influence the embedding space, leading to prototype drift and degraded performance on underrepresented or high-entropy domains. To address this, we introduce Cross-Discipline Variable Few-Shot Class-Incremental Learning (XD-VSCIL), a benchmark capturing real-world heterogeneity and imbalance where Domain Gravity naturally intensifies. We propose Hybrid Prototype Calibration (HyCal), a training-free method combining cosine similarity and Mahalanobis distance to capture complementary geometric properties-directional alignment and covariance-aware magnitude-yielding stable prototypes under imbalanced heterogeneous conditions. Operating on frozen CLIP embeddings, HyCal achieves consistent retention-adaptation improvements while maintaining efficiency. Experiments show HyCal effectively mitigates Domain Gravity and outperforms existing methods in imbalanced cross-domain incremental learning.

cs.CV

Easy to Learn, Yet Hard to Forget: Towards Robust Unlearning Under Bias

Machine unlearning, which enables a model to forget specific data, is crucial for ensuring data privacy and model reliability. However, its effectiveness can be severely undermined in real-world scenarios where models learn unintended biases from spurious correlations within the data. This paper investigates the unique challenges of unlearning from such biased models. We identify a novel phenomenon we term ``shortcut unlearning," where models exhibit an ``easy to learn, yet hard to forget" tendency. Specifically, models struggle to forget easily-learned, bias-aligned samples; instead of forgetting the class attribute, they unlearn the bias attribute, which can paradoxically improve accuracy on the class intended to be forgotten. To address this, we propose CUPID, a new unlearning framework inspired by the observation that samples with different biases exhibit distinct loss landscape sharpness. Our method first partitions the forget set into causal- and bias-approximated subsets based on sample sharpness, then disentangles model parameters into causal and bias pathways, and finally performs a targeted update by routing refined causal and bias gradients to their respective pathways. Extensive experiments on biased datasets including Waterbirds, BAR, and Biased NICO++ demonstrate that our method achieves state-of-the-art forgetting performance and effectively mitigates the shortcut unlearning problem.

cs.LG

Plug-in and Fine-tuning: Bridging the Gap between Small Language Models and Large Language Models

Large language models (LLMs) are renowned for their extensive linguistic knowledge and strong generalization capabilities, but their high computational demands make them unsuitable for resource-constrained environments. In contrast, small language models (SLMs) are computationally efficient but often lack the broad generalization capacity of LLMs. To bridge this gap, we propose PiFi, a novel framework that combines the strengths of both LLMs and SLMs to achieve high performance while maintaining efficiency. PiFi integrates a single frozen layer from an LLM into a SLM and fine-tunes the combined model for specific tasks, boosting performance without a significant increase in computational cost. We show that PiFi delivers consistent performance improvements across a range of natural language processing tasks, including both natural language understanding and generation. Moreover, our findings demonstrate PiFi's ability to effectively leverage LLM knowledge, enhancing generalization to unseen domains and facilitating the transfer of linguistic abilities.

cs.CL

Ultraslow water-mediated transmembrane interactions regulate the activation of A$_{\text{2A}}$ adenosine receptor

Water molecules inside G-protein coupled receptor have recently been spotlighted in a series of crystal structures. To decipher the dynamics and functional roles of internal waters in GPCR activity, we studied A$_{\text{2A}}$ adenosine receptor using $\mu$sec-molecular dynamics simulations. Our study finds that the amount of water flux across the transmembrane (TM) domain varies depending on the receptor state, and that the water molecules of the TM channel in the active state flow three times slower than those in the inactive state. Depending on the location in solvent-protein interface as well as the receptor state, the average residence time of water in each residue varies from $\sim\mathcal{O}(10^2)$ psec to $\sim\mathcal{O}(10^2)$ nsec. Especially, water molecules, exhibiting ultraslow relaxation ($\sim\mathcal{O}(10^2)$ nsec) in the active state, are found around the microswitch residues that are considered activity hotspots for GPCR function. A continuous allosteric network spanning the TM domain, arising from water-mediated contacts, is unique in the active state, underscoring the importance of slow waters in the GPCR activation.

q-bio.BM

Communication over the network of binary switches regulates the activation of A$_{2A}$ adenosine receptor

Dynamics and functions of G-protein coupled receptors (GPCRs) are accurately regulated by the type of ligands that bind to the orthosteric or allosteric binding sites. To glean the structural and dynamical origin of ligand-dependent modulation of GPCR activity, we performed total $\sim$ 5 $\mu$sec molecular dynamics simulations of A$_{2A}$ adenosine receptor (A$_{2A}$AR) in its apo, antagonist-bound, and agonist-bound forms in an explicit water and membrane environment, and examined the corresponding dynamics and correlation between the 10 key structural motifs that serve as the allosteric hotspots in intramolecular signaling network. We dubbed these 10 structural motifs "binary switches" as they display molecular interactions that switch between two distinct states. By projecting the receptor dynamics on these binary switches that yield $2^{10}$ microstates, we show that (i) the receptors in apo, antagonist-bound, and agonist-bound states explore vastly different conformational space; (ii) among the three receptor states the apo state explores the broadest range of microstates; (iii) in the presence of the agonist, the active conformation is maintained through coherent couplings among the binary switches; and (iv) to be most specific, our analysis shows that W246, located deep inside the binding cleft, can serve as both an agonist sensor and actuator of ensuing intramolecular signaling for the receptor activation.Finally, our analysis of multiple trajectories generated by inserting an agonist to the apo state underscores that the transition of the receptor from inactive to active form requires the disruption of ionic-lock in the DRY motif.

q-bio.BM

Mapping the intramolecular signal transduction of G-protein coupled receptors

G-protein coupled receptors (GPCRs), a major gatekeeper of extracellular signals on plasma membrane, are unarguably one of the most important therapeutic targets. Given the recent discoveries of allosteric modulations, an allosteric wiring diagram of intramolecular signal transductions would be of great use to glean the mechanism of receptor regulation. Here, by evaluating betweenness centrality ($C_B$) of each residue, we calculate maps of information flow in GPCRs and identify key residues for signal transductions and their pathways. Compared with preexisting approaches, the allosteric hotspots that our $C_B$-based analysis detects for A$_{2A}$ adenosine receptor (A$_{2A}$AR) and bovine rhodopsin are better correlated with biochemical data. In particular, our analysis outperforms other methods in locating the rotameric microswitches, which are generally deemed critical for mediating orthosteric signaling in class A GPCRs. For A$_{2A}$AR, the inter-residue cross-correlation map, calculated using equilibrium structural ensemble from molecular dynamics simulations, reveals that strong signals of long-range transmembrane communications exist only in the agonist-bound state. A seemingly subtle variation in structure, found in different GPCR subtypes or imparted by agonist bindings or a point mutation at an allosteric site, can lead to a drastic difference in the map of signaling pathways and protein activity. The signaling map of GPCRs provides valuable insights into allosteric modulations as well as reliable identifications of orthosteric signaling pathways.

q-bio.BM

Link between allosteric signal transduction and functional dynamics in a multi-subunit enzyme: S-adenosylhomocysteine hydrolase

S-adenosylhomocysteine hydrolase (SAHH), a cellular enzyme that plays a key role in methylation reactions including those required for maturation of viral mRNA, is an important drug target in the discovery of antiviral agents. While targeting the active site is a straightforward strategy of enzyme inhibition, evidences of allosteric modulation of active site in many enzymes underscore the molecular origin of signal transduction. Information of co-evolving sequences in SAHH family and the key residues for functional dynamics that can be identified using native topology of the enzyme provide glimpses into how the allosteric signaling network, dispersed over the molecular structure, coordinates intra- and inter-subunit conformational dynamics. To study the link between the allosteric communication and functional dynamics of SAHHs, we performed Brownian dynamics simulations by building a coarse-grained model based on the holo and ligand-bound structures. The simulations of ligand-induced transition revealed that the signal of intra-subunit closure dynamics is transmitted to form inter-subunit contacts, which in turn invoke a precise alignment of active site, followed by the dimer-dimer rotation that compacts the whole tetrameric structure. Further analyses of SAHH dynamics associated with ligand binding provided evidence of both induced fit and population shift mechanisms, and also showed that the transition state ensemble is akin to the ligand-bound state. Besides the formation of enzyme-ligand contacts at the active site, the allosteric couplings from the residues distal to the active site is vital to the enzymatic function.

q-bio.BM