arXiv ScienceSearch

arXiv subjects

Po Hu

Publications and source records attributed to Po Hu.

At least 19 recordsLinked to original sources

Generalized Category Discovery under Domain Shifts: From Vision to Vision-Language Models

Generalized Category Discovery (GCD) aims to categorize unlabelled instances from both known and unknown classes by transferring knowledge from labelled data of known classes. Existing methods assume all data comes from a single domain, yet real-world unlabelled data often exhibits domain shifts alongside semantic shifts. We study GCD under domain shifts and propose three frameworks that adapt foundation models, ranging from self-supervised vision models to vision-language models. (i) HiLo disentangles domain and semantic features through multi-level feature extraction and mutual information minimization, combined with PatchMix augmentation and curriculum sampling. (ii) HLPrompt extends HiLo with semantic-aware spatial prompt tuning to suppress background and domain noise. (iii) VLPrompt leverages vision-language models via factorized textual prompts and cross-modal consistency regularization. The three methods share core design principles while operating on different foundation backbones, making them suitable for different deployment scenarios. Extensive experiments on synthetic corruptions and real-world multi-domain shifts demonstrate consistent improvements over strong baselines. Project page: https://visual-ai.github.io/hilo/

cs.CV

Debate to Align: Reliable Entity Alignment through Two-Stage Multi-Agent Debate

Entity alignment (EA) aims to identify entities referring to the same real-world object across different knowledge graphs (KGs). Recent approaches based on large language models (LLMs) typically obtain entity embeddings through knowledge representation learning and use embedding similarity to identify an alignment-uncertain entity set. For each uncertain entity, a candidate entity set (CES) is then retrieved based on embedding similarity to support subsequent alignment reasoning and decision making. However, the reliability of the CES and the reasoning capability of LLMs critically affect the effectiveness of subsequent alignment decisions. To address this issue, we propose AgentEA, a reliable EA framework based on multi-agent debate. AgentEA first improves embedding quality through entity representation preference optimization, and then introduces a two-stage multi-role debate mechanism consisting of lightweight debate verification and deep debate alignment to progressively enhance the reliability of alignment decisions while enabling more efficient debate-based reasoning. Extensive experiments on public benchmarks under cross-lingual, sparse, large-scale, and heterogeneous settings demonstrate the effectiveness of AgentEA.

cs.CL

DSCD: Large Language Model Detoxification with Self-Constrained Decoding

Detoxification in large language models (LLMs) remains a significant research challenge. Existing decoding detoxification methods are all based on external constraints, which require additional resource overhead and lose generation fluency. This work proposes Detoxification with Self-Constrained Decoding (DSCD), a novel method for LLM detoxification without parameter fine-tuning. DSCD strengthens the inner next-token distribution of the safety layer while weakening that of hallucination and toxic layers during output generation. This effectively diminishes toxicity and enhances output safety. DSCD offers lightweight, high compatibility, and plug-and-play capabilities, readily integrating with existing detoxification methods for further performance improvement. Extensive experiments on representative open-source LLMs and public datasets validate DSCD's effectiveness, demonstrating state-of-the-art (SOTA) performance in both detoxification and generation fluency, with superior efficiency compared to existing methods. These results highlight DSCD's potential as a practical and scalable solution for safer LLM deployments.

cs.CL

FediLoRA: Practical Federated Fine-Tuning of Foundation Models Under Missing-Modality Constraints

Federated Learning with LoRA fine-tuning offers an efficient and privacy-aware solution for institutions to collaboratively leverage their large datasets to train VLLMs. However, participating institutions often possess heterogeneous computational resources, resulting in imbalanced LoRA ranks, which pose a major challenge for effective collaboration. In addition, real-world applications in domains such as healthcare and transportation frequently suffer from missing modalities due to user mistakes or device failures, which significantly degrade global model performance in federated settings. To the best of our knowledge, no prior work has addressed these two challenges simultaneously in federated VLLMs. To tackle these issues, we propose FediLoRA, a lightweight federated LoRA aggregation framework that effectively mitigates the impact of missing modalities in heterogeneous environment. FediLoRA is explicitly motivated by the observation that simple averaging and structured editing can jointly benefit both global and personalized models. Our approach achieves strong performance across multiple general-domain and medical-domain benchmark datasets. Additional experiments on healthcare data further demonstrate that FediLoRA is well-suited for practical, real-world deployment scenarios. Our code is released at https://github.com/gotobcn8/FediLoRA.

cs.LG

Equivariant operations in topological Hochschild homology

We observe a new equivariant relationship between topological Hochschild homology and cohomology. We also calculate the topological Hochschild homology of the topological Hochschild cohomology of a finite prime field, which can be viewed as a certain ring of structured operations in this case.

math.AT

The $\mathbb{Z}/p$-equivariant spectrum $BP\mathbb{R}$ for an odd prime $p$

In the present paper, we construct a $\mathbb{Z}/p$-equivariant analog of the $\mathbb{Z}/2$-equivariant spectrum $BP\mathbb{R}$ previously constructed by Hu and Kriz. We prove that this spectrum has some of the properties conjectured by Hill, Hopkins, and Ravenel. Our main construction method is an $\mathbb{Z}/p$-equivariant analog of the Brown-Peterson tower of $BP$, based on a previous description of the $\mathbb{Z}/p$-equivariant Steenrod algebra with constant coefficients by the authors. We also describe several variants of our construction and comparisons with other known equivariant spectra.

math.AT

Graph-augmented Learning to Rank for Querying Large-scale Knowledge Graph

Knowledge graph question answering (KGQA) based on information retrieval aims to answer a question by retrieving answer from a large-scale knowledge graph. Most existing methods first roughly retrieve the knowledge subgraphs (KSG) that may contain candidate answer, and then search for the exact answer in the KSG. However, the KSG may contain thousands of candidate nodes since the knowledge graph involved in querying is often of large scale, thus decreasing the performance of answer selection. To tackle this problem, we first propose to partition the retrieved KSG to several smaller sub-KSGs via a new subgraph partition algorithm and then present a graph-augmented learning to rank model to select the top-ranked sub-KSGs from them. Our proposed model combines a novel subgraph matching networks to capture global interactions in both question and subgraphs, and an Enhanced Bilateral Multi-Perspective Matching model is proposed to capture local interactions. Finally, we apply an answer selection model on the full KSG and the top-ranked sub-KSGs respectively to validate the effectiveness of our proposed graph-augmented learning to rank method. The experimental results on multiple benchmark datasets have demonstrated the effectiveness of our approach.

cs.CL

Triples-to-Text Generation with Reinforcement Learning Based Graph-augmented Neural Networks

Considering a collection of RDF triples, the RDF-to-text generation task aims to generate a text description. Most previous methods solve this task using a sequence-to-sequence model or using a graph-based model to encode RDF triples and to generate a text sequence. Nevertheless, these approaches fail to clearly model the local and global structural information between and within RDF triples. Moreover, the previous methods also face the non-negligible problem of low faithfulness of the generated text, which seriously affects the overall performance of these models. To solve these problems, we propose a model combining two new graph-augmented structural neural encoders to jointly learn both local and global structural information in the input RDF triples. To further improve text faithfulness, we innovatively introduce a reinforcement learning (RL) reward based on information extraction (IE). We first extract triples from the generated text using a pretrained IE model and regard the correct number of the extracted triples as the additional RL reward. Experimental results on two benchmark datasets demonstrate that our proposed model outperforms the state-of-the-art baselines, and the additional reinforcement learning reward does help to improve the faithfulness of the generated text.

cs.CL

Coefficients of the $\Sigma_3$-equivariant complex cobordism ring

In this paper, we calculate the coefficient ring of equivariant Thom complex cobordism for the symmetric group on three elements. We also make some remarks on general methods of calculating certain pullbacks of rings which typically occur in calculations of equivariant cobordism.

math.AT

Development of Gated Fiber Detectors for Laser-Induced Strong Electromagnetic Pulse Environments

With the development of laser technologies, nuclear reactions can happen in high-temperature plasma environments induced by lasers and have attracted a lot of attention from different physical disciplines. However, studies on nuclear reactions in plasma are still limited by detecting technologies. This is mainly due to the fact that extremely high electromagnetic pulses (EMPs) can also be induced when high-intensity lasers hit targets to induce plasma, and then cause dysfunction of many types of traditional detectors. Therefore, new particle detecting technologies are highly needed. In this paper, we report a recently developed gated fiber detector which can be used in harsh EMP environments. In this prototype detector, scintillating photons are coupled by fiber and then transferred to a gated photomultiplier tube which is located far away from the EMP source and shielded well. With those measures, the EMPs can be avoided, and this device has the capability to identify a single event of nuclear reaction products generated in laser-induced plasma from noise EMP backgrounds. This new type of detector can be widely used as a Time-of-Flight (TOF) detector in high-intensity laser nuclear physics experiments for detecting neutron, photons, and other charged particles.

physics.ins-det

Tate cohomology of connected k-theory for elementary abelian groups revisited

Tate cohomology (as well as Borel homology and cohomology) of connective K-theory for $G=(\mathbb{Z}/2)^n$ was completely calculated by Bruner and Greenlees. In this note, we essentially redo the calculation by a different, more elementary method, and we extend it to $p>2$ prime. We also identify the resulting spectra, which are products of Eilenberg-Mac Lane spectra, and finitely many finite Postnikov towers. For $p=2$, we also reconcile our answer completely with the result of Bruner and Greenlees, which is in a different form, and hence the comparison involves some non-trivial combinatorics.

math.KT

Derived representation theory of Lie algebras and stable homotopy categorification of $sl_k$

We set up foundations of representation theory over $S$, the sphere spectrum, which is the `initial ring' of stable homotopy theory. In particular, we treat $S$-Lie algebras and their representations, characters, $gl_n(S)$-Verma modules and their duals, Harish-Chandra pairs and Zuckermann functors. As an application, we construct a Khovanov $sl_k$-stable homotopy type with a large prime hypothesis, which is a new link invariant, using a stable homotopy analogue of the method of J.Sussan.

math.AT

On some adjunctions in equivariant stable homotopy theory

We investigate certain adjunctions in derived categories of equivariant spectra, including a right adjoint to fixed points, a right adjoint to pullback by an isometry of universes, and a chain of two right adjoints to geometric fixed points. This leads to a variety of interesting other adjunctions, including a chain of 6 (sometimes 7) adjoints involving the restriction functor to a subgroup of a finite group on equivariant spectra indexed over the trivial universe.

math.AT

D-structures and derived Koszul duality for unital operad algebras

Generalizing a concept of Lipshitz, Ozsv\'ath and Thurs-ton from Bordered Floer homology, we define $D$-structures on algebras of unital operads, which can also be interpreted as a generalization of a seemingly unrelated concept of Getzler and Jones. This construction gives rise to an equivalence of derived categories, which can be thought of as a unital version of Koszul duality using non-unital Quillen homology. We also discuss a multi-sorted version of the construction, which provides a framework for unifying the known algebraic contexts of Koszul duality.

math.KT

Equivariant K-theory of compact Lie groups with involution

For a compact simply connected simple Lie group $G$ with an involution $\alpha$, we compute the $G\rtimes \Z/2$-equivariant K-theory of $G$ where $G$ acts by conjugation and $\Z/2$ acts either by $\alpha$ or by $g\mapsto \alpha(g)^{-1}$. We also give a representation-theoretic interpretation of those groups, as well as of $K_G(G)$.

math.KT