arXiv ScienceSearch

arXiv subjects

Chang Han

Publications and source records attributed to Chang Han.

14 recordsLinked to original sources

MOTIF: Motivation-guided Topology Inference for Cold-start Multimodal Recommendation

Cold-start multimodal recommendation faces three coupled challenges: (i) sparse interactions obscure user intent, (ii) cold items remain topologically isolated, and (iii) similarity-based item graphs may cause semantic drift. To address these issues, we propose MOTIF, a Motivation-guided Topology Inference framework for cold-start multimodal recommendation. MOTIF integrates Semantic Motivation Reasoning, Knowledge-enhanced Graph Reconstruction, Weighted Graph Contrastive Learning, and Semantic-Structural Alignment. It uses offline LLM reasoning to infer motivation semantics, reconstructs transferable item-item topology, and learns robust graph embeddings without injecting generated text into prediction. Experiments on three multimodal benchmarks show consistent gains over graph-based, multimodal, cold-start, and LLM-enhanced baselines, with up to 6.07% relative improvement over the strongest recent baseline.

cs.IR

On Defining Chart Types Boundaries

What makes a Gantt chart? This question proved unexpectedly difficult to answer when we set out to build a design space for Gantt charts. Existing definitions, each shaped by their respective research goals, made different scope choices that we could not directly reconcile. We reasoned about what should and should not count as a Gantt chart, developing concepts and tools along the way. We distinguish features that are essential to a chart type's identity from those that can vary, and use these distinctions to map how chart types relate through what they share and lack. Applying these ideas to Gantt charts, radar charts, and table cartograms, we produce key insights on what boundary work reveals: definitions diverge for functional reasons, drawing boundaries exposes hidden structure in descriptive vocabulary such as feature entanglements, and scope choices shape how far findings can generalize. We came to understand that there is not a definitive answer, but that working through the question produced a functional definition that guided the design space we originally set out to build. Additionally, we present vocabulary and tools for reasoning about chart type boundaries and surfacing these boundary decisions, alongside a documented Gantt chart design space. Our broader reflection is that scope choices in chart-type-centered research---which determine what design spaces include, what grammars generate, and what perceptual studies measure---are research decisions worth making visible.

cs.HC

Primordial Black Holes from Vector-Induced Curvature Perturbations Sourced by Primordial Magnetic Fields

Generating an appreciable abundance of primordial black holes (PBHs) requires a substantial enhancement of primordial curvature perturbations on small scales. In this work, we propose a new post-inflationary mechanism in which such an enhancement arises during a stiff, or kination, epoch. The mechanism is driven by metric vector perturbations sourced by the vector component of the electromagnetic stress-energy tensor associated with primordial magnetic fields (PMFs). Since these first-order vector modes remain approximately constant during kination, they act as persistent nonlinear sources for second-order scalar perturbations. We show that the resulting vector-induced curvature perturbations are amplified toward the infrared cutoff of the kination band and exhibit the characteristic scaling $\mathcal P_{\mathcal R}(k)\propto k^{-5}$. As a concrete realization, we consider PMFs generated in a Ratra-type magnetogenesis scenario and find that the induced curvature perturbations can produce PBHs with an abundance large enough to constitute a substantial fraction of the dark matter.

gr-qc

EAGLE-Pangu: Accelerator-Safe Tree Speculative Decoding on Ascend NPUs

Autoregressive decoding remains a primary bottleneck in large language model (LLM) serving, motivating speculative decoding methods that reduce expensive teacher-model invocations by verifying multiple candidate tokens per step. Tree-structured speculation further increases parallelism, but is often brittle when ported across heterogeneous backends and accelerator stacks, where attention masking, KV-cache layouts, and indexing semantics are not interchangeable. We present EAGLE-Pangu, a reproducible system that ports EAGLE-3-style tree speculative decoding to a Pangu teacher backend on Ascend NPUs. EAGLE-Pangu contributes (i) an explicit branch/commit cache manager built on the Cache API, (ii) accelerator-safe tree tensorization that removes undefined negative indices by construction and validates structural invariants, and (iii) a fused-kernel-compatible teacher verification path with a debuggable eager fallback. On 240 turns from MT-Bench and HumanEval-style prompts, EAGLE-Pangu improves end-to-end decoding throughput by 1.27x on average, up to 2.46x at p99, over teacher-only greedy decoding in the fused-kernel performance path. We also provide a fused-kernel-free reference path with structured traces and invariant checks to support reproducible debugging and ablation across execution modes and tree budgets.

cs.LG

The Feng Shui of Visualization: Design the Path to SUCCESS and GOOD FORTUNE

Superstition and religious belief system have historically shaped human behavior, offering powerful psychological motivations and persuasive frameworks to guide actions. Inspired by Feng Shui -- an ancient Chinese superstition -- this paper proposes a pseudo-theoretical framework that integrates superstition-like heuristics into visualization design. Rather than seeking empirical truth, this framework leverages culturally resonant (superstitious) narratives and symbolic metaphors as persuasive tools to encourage desirable design practices, such as clarity, accessibility, and audience-centered thinking. We articulate a set of visualization designs into a Feng Shui compass, reframing empirical design principles and guidelines within an engaing mythology. We present how visualization design principles can be intepreted in Feng Shui narratives, discussing the potential of these metaphorical principles in reducing designer anxiety, fostering community norms, and enhancing the memorability and internalization of visualization design guidelines. Finally, we discuss Feng Shui visualization theory as a set of cognitive shortcuts that can exert persuasive power through playful, belief-like activities.

cs.HC

FedSODA: Federated Fine-tuning of LLMs via Similarity Group Pruning and Orchestrated Distillation Alignment

Federated fine-tuning (FFT) of large language models (LLMs) has recently emerged as a promising solution to enable domain-specific adaptation while preserving data privacy. Despite its benefits, FFT on resource-constrained clients relies on the high computational and memory demands of full-model fine-tuning, which limits the potential advancement. This paper presents FedSODA, a resource-efficient FFT framework that enables clients to adapt LLMs without accessing or storing the full model. Specifically, we first propose a similarity group pruning (SGP) module, which prunes redundant layers from the full LLM while retaining the most critical layers to preserve the model performance. Moreover, we introduce an orchestrated distillation alignment (ODA) module to reduce gradient divergence between the sub-LLM and the full LLM during FFT. Through the use of the QLoRA, clients only need to deploy quantized sub-LLMs and fine-tune lightweight adapters, significantly reducing local resource requirements. We conduct extensive experiments on three open-source LLMs across a variety of downstream tasks. The experimental results demonstrate that FedSODA reduces communication overhead by an average of 70.6%, decreases storage usage by 75.6%, and improves task accuracy by 3.1%, making it highly suitable for practical FFT applications under resource constraints.

cs.LG

Teaching Critical Visualization: A Field Report

Critical Visualization is gaining popularity and academic focus, yet relatively few academic courses have been offered to support students in this complex area. This experience report describes a recent experimental course on the topic, exploring both what the topic could be as well as an experimental content structure (namely as scavenger hunt). Generally the course was successful, achieving the learning objectives of developing critical thinking skills, improving communication about complex ideas, and developing a knowledge about theories in the area. While improvements can be made, we hope that humanistic notions of criticality are embraced more deeply in visualization pedagogy.

cs.HC

Infrared Behavior of Induced Gravitational Waves from Isocurvature Perturbations

Induced gravitational waves provide a powerful probe of primordial perturbations in the early universe through their distinctive spectral properties. We analyze the spectral energy density $\Omega_{\text{GW}}$ of gravitational waves induced by isocurvature scalar perturbations. In the infrared regime, we find that the spectral slope $n_{\text{GW}} \equiv \text{d} \ln\Omega_\mathrm{GW}/\text{d}\ln k$ takes the log-dependent form $3-4/ \ln (\tilde{k}_*^2 / 6k^2)$, where $\tilde{k}_*$ represents the effective peak scale of the primordial scalar power spectrum. This characteristic behavior differs markedly from that of adiabatic-induced gravitational waves, establishing a robust observational discriminant between isocurvature and adiabatic primordial perturbation modes.

gr-qc

Constraining inflation with nonminimal derivative coupling with the Parkes Pulsar Timing Array third data release

We study an inflation model with nonminimal derivative coupling that features a coupling between the derivative of the inflaton field and the Einstein tensor. This model naturally amplifies curvature perturbations at small scales via gravitationally enhanced friction, a mechanism critical for the formation of primordial black holes and the associated production of potentially detectable scalar-induced gravitational waves. We derive analytical expressions for the primordial power spectrum, enabling efficient exploration of the model parameter space without requiring computationally intensive numerical solutions of the Mukhanov-Sasaki equation. Using the third data release of the Parkes Pulsar Timing Array (PPTA DR3), we constrain the model parameters characterizing the coupling function: $\phi_c = 3.7^{+0.3}_{-0.5} M_\mathrm{P}$, $\log_{10} \omega_L = 7.1^{+0.6}_{-0.3}$, and $\log_{10} \sigma = -8.3^{+0.3}_{-0.6}$ at 90\% confidence level. Our results demonstrate the growing capability of pulsar timing arrays to probe early Universe physics, complementing traditional cosmic microwave background observations by providing unique constraints on inflationary dynamics at small scales.

gr-qc

A Deixis-Centered Approach for Documenting Remote Synchronous Communication around Data Visualizations

Referential gestures, or as termed in linguistics, deixis, are an essential part of communication around data visualizations. Despite their importance, such gestures are often overlooked when documenting data analysis meetings. Transcripts, for instance, fail to capture gestures, and video recordings may not adequately capture or emphasize them. We introduce a novel method for documenting collaborative data meetings that treats deixis as a first-class citizen. Our proposed framework captures cursor-based gestural data along with audio and converts them into interactive documents. The framework leverages a large language model to identify word correspondences with gestures. These identified references are used to create context-based annotations in the resulting interactive document. We assess the effectiveness of our proposed method through a user study, finding that participants preferred our automated interactive documentation over recordings, transcripts, and manual note-taking. Furthermore, we derive a preliminary taxonomy of cursor-based deictic gestures from participant actions during the study. This taxonomy offers further opportunities for better utilizing cursor-based deixis in collaborative data analysis scenarios.

cs.HC

An Overview + Detail Layout for Visualizing Compound Graphs

Compound graphs are networks in which vertices can be grouped into larger subsets, with these subsets capable of further grouping, resulting in a nesting that can be many levels deep. In several applications, including biological workflows, chemical equations, and computational data flow analysis, these graphs often exhibit a tree-like nesting structure, where sibling clusters are disjoint. Common compound graph layouts prioritize the lowest level of the grouping, down to the individual ungrouped vertices, which can make the higher level grouped structures more difficult to discern, especially in deeply nested networks. Leveraging the additional structure of the tree-like nesting, we contribute an overview+detail layout for this class of compound graphs that preserves the saliency of the higher level network structure when groups are expanded to show internal nested structure. Our layout draws inner structures adjacent to their parents, using a modified tree layout to place substructures. We describe our algorithm and then present case studies demonstrating the layout's utility to a domain expert working on data flow analysis. Finally, we discuss network parameters and analysis situations in which our layout is well suited.

cs.HC

Exploring the Interactions between Target Positive and Negative Information for Acoustic Echo Cancellation

Acoustic echo cancellation (AEC) aims to remove interference signals while leaving near-end speech least distorted. As the indistinguishable patterns between near-end speech and interference signals, near-end speech can't be separated completely, causing speech distortion and interference signals residual. We observe that besides target positive information, e.g., ground-truth speech and features, the target negative information, such as interference signals and features, helps make pattern of target speech and interference signals more discriminative. Therefore, we present a novel AEC model encoder-decoder architecture with the guidance of negative information termed as CMNet. A collaboration module (CM) is designed to establish the correlation between the target positive and negative information in a learnable manner via three blocks: target positive, target negative, and interactive block. Experimental results demonstrate our CMNet achieves superior performance than recent methods.

eess.AS

A study of the limits of imaging capability due to water scattering effects in underwater ghost imaging

Underwater ghost imaging is an effective means of underwater detection. In this paper, a theoretical and experimental study of underwater ghost imaging is carried out by combining the description of underwater optical field transmission with the inherent optical parameters of the water body. This paper utilizes the Wells model and the approximate S-S scattering phase function to create a model for optical transmission underwater. The second-order Glauber function of the optical field is then employed to analyze the scattering field's degradation during the transmission process. This analysis is used to evaluate the impact of the water body on ghost imaging. The simulation and experimental results verify that the proposed underwater ghost imaging model can better describe the degradation effect of water bodies on ghost imaging. A series of experiments comparing underwater ghost imaging at different detection distances are also carried out in this paper. In the experiments, cooperative targets can be imaged up to 65.2m (9.3AL, at attenuation coefficient c=0.1426m-1 and the scattering coefficient b=0.052m-1) and non-cooperative targets up to 41.2m (6.4AL, at c=0.1569m-1 and b=0.081m-1) . By equating the experimental maximum imaged attenuation length for cooperative targets to Jerlov-I water (b=0.002m-1 and a=0.046m-1), the system will have a maximum imaging distance of 193m. Underwater ghost imaging is expected to achieve longer-range imaging by optimizing the system emission energy and detection sensitivity.

physics.optics

All Information is Necessary: Integrating Speech Positive and Negative Information by Contrastive Learning for Speech Enhancement

Monaural speech enhancement (SE) is an ill-posed problem due to the irreversible degradation process. Recent methods to achieve SE tasks rely solely on positive information, e.g., ground-truth speech and speech-relevant features. Different from the above, we observe that the negative information, such as original speech mixture and speech-irrelevant features, are valuable to guide the SE model training procedure. In this study, we propose a SE model that integrates both speech positive and negative information for improving SE performance by adopting contrastive learning, in which two innovations have consisted. (1) We design a collaboration module (CM), which contains two parts, contrastive attention for separating relevant and irrelevant features via contrastive learning and interactive attention for establishing the correlation between both speech features in a learnable and self-adaptive manner. (2) We propose a contrastive regularization (CR) built upon contrastive learning to ensure that the estimated speech is pulled closer to the clean speech and pushed far away from the noisy speech in the representation space by integrating self-supervised models. We term the proposed SE network with CM and CR as CMCR-Net. Experimental results demonstrate that our CMCR-Net achieves comparable and superior performance to recent approaches.

eess.AS