arXiv ScienceSearch

arXiv subjects

Pin-Han Ho

Publications and source records attributed to Pin-Han Ho.

12 recordsLinked to original sources

Sensor-Conditioned Representation Learning via Scene-Relevant Observation Quotients

Learned representations in intelligent sensing systems are often evaluated by reconstruction fidelity or downstream prediction accuracy, but these criteria do not specify which latent distinctions are justified by the sensing process. In sensor-conditioned environments, nuisance factors can change measurements without changing the scene, while distinct scenes may be indistinguishable under limited sensing capability. This paper formulates sensor-conditioned representation correctness as preserving sensing-supported scene distinctions while suppressing nuisance-induced and sensor-unsupported variation. We introduce the scene-relevant observation quotient, a representation target induced by sensing-supported distinguishability after nuisance canonicalization, and develop Observation-Quotient Tucker-Structured Autoencoding (OQ-TSAE), a scene-nuisance factorized framework with diagnostics for false distinction, false merge, nuisance sensitivity, and latent ordering consistency. Experiments on a controlled benchmark show that quotient-consistent supervision improves representation-correctness diagnostics over reconstruction-oriented, metric-learning, and contrastive-learning baselines. Sensitivity, perturbation, and ablation studies show the importance of quotient-aligned supervision, reliable quotient relations, and quotient geometry. Complementary real-radar experiments show that a reconstruction-only OQ-TSAE variant retains competitive downstream utility, robustness under observation degradation, and low seed-to-seed variability. These results suggest that sensor-conditioned representations should be evaluated not only by predictive utility, but also by whether their latent geometry preserves sensing-justified scene distinctions.

cs.AI

Query-Conditioned Knowledge Alignment for Reliable Cross-System Medical Reasoning

Cross-domain knowledge alignment is essential for integrating heterogeneous medical systems, yet existing approaches typically treat entity alignment as a static matching problem, ignoring query context and cross-system asymmetry. This limitation is particularly critical in integrative medical settings, where correspondence between concepts is inherently context-dependent, non-bijective, and direction-sensitive. In this paper, we propose Query-Conditioned Entity Alignment (QCEA), which reformulates entity alignment as a query-conditioned correspondence problem. Instead of learning a fixed mapping between entity representations, QCEA treats the textual description of a source entity as a query and ranks candidate entities in the target graph, enabling context-dependent alignment. The framework integrates semantic encoding, graph-based representation learning, and a direction-aware transformation module to capture asymmetric and many-to-many correspondence across heterogeneous knowledge systems. We evaluate QCEA on TCM--WM knowledge graphs derived from SymMap, covering both symptom alignment and herb--molecule alignment tasks. Experimental results show consistent improvements over representative baselines, particularly on rank-sensitive metrics such as Hit@K and MRR. Furthermore, downstream retrieval-augmented generation (RAG) experiments demonstrate that improved alignment leads to better evidence retrieval, stronger grounding, and higher answer accuracy. These findings highlight that alignment is not merely a data integration step, but a key factor that shapes knowledge accessibility and reliability in cross-system medical reasoning.

cs.AI

Geometrical Cross-Attention and Nonvoid Voxelization for Efficient 3D Medical Image Segmentation

Accurate segmentation of 3D medical scans is crucial for clinical diagnostics and treatment planning, yet existing methods often fail to achieve both high accuracy and computational efficiency across diverse anatomies and imaging modalities. To address these challenges, we propose GCNV-Net, a novel 3D medical segmentation framework that integrates a Tri-directional Dynamic Nonvoid Voxel Transformer (3DNVT), a Geometrical Cross-Attention module (GCA), and Nonvoid Voxelization. The 3DNVT dynamically partitions relevant voxels along the three orthogonal anatomical planes, namely the transverse, sagittal, and coronal planes, enabling effective modeling of complex 3D spatial dependencies. The GCA mechanism explicitly incorporates geometric positional information during multi-scale feature fusion, significantly enhancing fine-grained anatomical segmentation accuracy. Meanwhile, Nonvoid Voxelization processes only informative regions, greatly reducing redundant computation without compromising segmentation quality, and achieves a 56.13% reduction in FLOPs and a 68.49% reduction in inference latency compared to conventional voxelization. We evaluate GCNV-Net on multiple widely used benchmarks: BraTS2021, ACDC, MSD Prostate, MSD Pancreas, and AMOS2022. Our method achieves state-of-the-art segmentation performance across all datasets, outperforming the best existing methods by 0.65% on Dice, 0.63% on IoU, 1% on NSD, and relatively 14.5% on HD95. All results demonstrate that GCNV-Net effectively balances accuracy and efficiency, and its robustness across diverse organs, disease conditions, and imaging modalities highlights strong potential for clinical deployment.

cs.CV

Low-Complexity Algorithm for Stackelberg Prediction Games with Global Optimality

Stackelberg prediction games (SPGs) model strategic data manipulation in adversarial learning via a leader--follower interaction between a learner and a self-interested data provider, leading to challenging bilevel optimization problems. Focusing on the least-squares setting (SPG-LS), recent work shows that the bilevel program admits an equivalent spherically constrained least-squares (SCLS) reformulation, which avoids costly conic programming and enables scalable algorithms. In this paper, we develop a simple and efficient alternating direction method of multiplier (ADMM) based solver for the SCLS problem. By introducing a consensus splitting that separates the quadratic objective from the spherical constraint, we obtain an augmented Lagrangian formulation with closed-form updates: the primal quadratic step reduces to solving a fixed shifted linear system, the constraint step is a projection onto the unit sphere, and the dual step is a lightweight scaled ascent. The resulting method has low per-iteration complexity and allows pre-factorization of the constant system matrix for substantial speedups. Experiments demonstrate that the proposed ADMM approach achieves competitive solution quality with significantly improved computational efficiency compared with existing global solvers for SCLS, particularly in sparse and high-dimensional regimes.

eess.SP

Observability Engineering: From Measurement to Information Generation in Active Sensing Systems

This article develops a four stage observability engineering framework for interaction driven sensing. The framework integrates physical modeling, information geometry, symmetry reduction, cross system correspondence, and observable world models into a unified engineering perspective, where sensing actions induce information structures that govern local distinguishability, while physical symmetries define equivalence classes of states that remain fundamentally indistinguishable. The proposed framework explains several common phenomena in modern sensing systems. Physically different architectures can exhibit comparable sensing capability when they induce similar quotient space information structures, whereas insufficient action diversity can lead to blind directions, ill conditioned estimation, and false identifiability. The framework also provides a quotient space language for comparing heterogeneous sensing architectures beyond hardware level descriptors, while clarifying how different forms of sensing diversity contribute to distinguishability. The framework is developed primarily through high frequency electromagnetic sensing, where geometric optics approximations provide an analytically interpretable physical realization. The broader observability engineering perspective, however, does not depend on geometric optics itself; it requires an action indexed observation family and a task relevant distinguishability measure. From this viewpoint, observability is not a passive property of sensor measurements alone, but an engineered outcome of action physics coupling under physical and symmetry constraints.

cs.ET

Knowledization: Claim-Level Epistemic Control with Admissibility, Source Support, and Real-World Proxy Boundaries

Near field mmWave sensing is poised to play a key role in future wireless systems, enabling environment-aware, embodied, and application adaptive operation under stringent form-factor and hardware constraints. However, achieving high spatial resolution in the near field typically requires large antenna arrays, multiple radio frequency (RF) chains, or mechanical scanning, creating a fundamental tension between spatial observability and system simplicity. This paper presents frequency as aperture clip on antenna fabric (FaACAF), a hardware efficient sensing by design architecture that synthesizes spatial aperture through the FaA paradigm using a single RF chain. FaACAF realizes a modular clip on aperture fabric, in which frequency selective clip on modules (CMs) are attached to a shared guided-wave substrate and implicitly coordinated by the instantaneous frequency modulated continuous wave (FMCW) excitation frequency. In this fabric, FMCW signaling simultaneously indexes the sensing aperture and orchestrates uplink/downlink signal distribution and echo multiplexing in a switch free, fully passive, and all analog manner, eliminating RF switching and multichannel front ends. An online self calibration mechanism stabilizes the frequency to aperture mapping under practical attachment variability without requiring full matrix calibration. Two case studies illustrate the robustness of the proposed approach and quantify the predictable sensing margin tradeoffs introduced by modular deployment. Overall, FaACAF demonstrates that near field spatial observability can be scaled through architectural coordination in the frequency domain rather than hardware expansion, providing a reconfigurable and hardware efficient pathway toward embodied sensing and integrated sensing and communication (ISAC) in future wireless systems.

cs.IT

Task-Aware Identifiability: Observation Quotients, Statistical Geometry, and Representation Accessibility

Intelligent systems often assume observations contain the information required for a task. This fails struc- turally when a declared mechanism assigns identical ob- servation laws to latent states the task must distin- guish. We develop a unified diagnostic framework con- necting observation-induced equivalence, task identifia- bility, local statistical geometry, representation accessi- bility, and achievable risk. Its organizing condition is a standard factorization property: the observation parti- tion refines the task partition, and conditionally iid rep- etitions of that fixed channel cannot recover an exactly collapsed distinction. On regular quotient strata, Fisher geometry and nuisance-efficient Fisher information char- acterize first-order local resolvability. For representa- tions, we distinguish law-valued encoder experiments from deterministic quotient embeddings and exact task sufficiency from scale-controlled decoder accessibility. We organize task risk into structural, finite-observation, representation-channel, and learning gaps. Controlled synthetic studies validate these layers; a physics-based synthetic near-field model yields sharply different range and angle information, while NYUv2 shows persistent, decoder-modulated task preferences among frozen en- coders under two pointwise probes. TAI thereby local- izes whether task-relevant information is absent at ob- servation, weakly resolved, lost or poorly accessible in representation, or left to the learner.

cs.AR

Epistemic Memory: A Validity Layer for Self-Maintaining Intelligent Systems

AI memory mechanisms primarily focus on preserving information content, often neglecting the validity conditions under which knowledge remains applicable, leading to semantic coordinate drift when agents move, change sensors, or encounter novel environments. This paper proposes epistemic memory as a validity-maintenance layer that governs when stored knowledge remains applicable. We formalize the dynamic epistemic quotient, an observation-induced equivalence structure over hypotheses that evolves with an agent's interaction history and sensing capabilities. We derive a pairwise incompatibility lower bound showing that any fixed semantic representation must incur unavoidable error across changing epistemic boundaries. We further identify a failure mode under symmetric crossing quotients, where overlap-based class-level transport becomes independent of pre-transition belief, motivating preservation of within-class hypothesis provenance. We introduce Observable Belief Memory (OBM) as a constructive epistemic governance architecture, combining the current epistemic quotient, belief over quotient classes, and within-class provenance through a posterior-consistent update. Controlled experiments demonstrate that explicit epistemic tracking improves robustness under changing observation conditions, particularly when class-level correspondence becomes insufficient. These results suggest that self-maintaining intelligent systems may benefit from reasoning not only about hidden states, but also about the evolving validity of the representations through which those states become knowable.

cs.LG

On Achieving High-Fidelity Grant-free Non-Orthogonal Multiple Access

Grant-free access (GFA) has been envisioned to play an active role in massive Machine Type Communication (mMTC) under 5G and Beyond mobile systems, which targets at achieving significant reduction of signaling overhead and access latency in the presence of sporadic traffic and small-size data. The paper focuses on a novel K-repetition GFA (K-GFA) scheme by incorporating Reed-Solomon (RS) code with the contention resolution diversity slotted ALOHA (CRDSA), aiming to achieve high-reliability and low-latency access in the presence of massive uncoordinated MTC devices (MTCDs). We firstly defines a MAC layer transmission structure at each MTCD for supporting message-level RS coding on a data message of $Q$ packets, where a RS code of $KQ$ packets is generated and sent in a super time frame (STF) that is composed of $Q$ time frames. The access point (AP) can recover the original $Q$ packets of the data message if at least $Q$ out of the $KQ$ packets of the RS code are successfully received. The AP buffers the received MTCD signals of each resource block (RB) within an STF and exercises the CRDSA based multi-user detection (MUD) by exploring signal-level inter-RB correlation via iterative interference cancellation (IIC). With the proposed CRDSA based K-GFA scheme, we provide the complexity analysis, and derive a closed-form analytical model on the access probability for each MTCD as well as its simplified approximate form. Extensive numerical experiments are conducted to validate its effectiveness on the proposed CRDSA based K-GFA scheme and gain deep understanding on its performance regarding various key operational parameters.

cs.IT

A Novel Framework of K-repetition Grant-free Access via Diversity Slotted Aloha (DSA)

This article introduces a novel framework of multi-user detection (MUD) for K-repetition grant-free non-orthogonal multiple access (K-GF-NOMA), called $\alpha$ iterative interference cancellation diversity slotted aloha ($\alpha$-IIC-DSA). The proposed framework targets at a simple yet effective decoding process where the AP can intelligently exploit the correlation among signals received at different resource blocks (RBs) so as to generate required multi-access interference (MAI) for realizing the signal-interference cancellation (SIC) based MUD. By keeping all operation and hardware complexity at the access point (AP), the proposed framework is applicable to the scenarios with random and uncoordinated access by numerous miniature mMTC devices (MTCDs). Numerical experiments are conducted to gain deep understanding on the performance of launching the proposed framework for K-GF-NOMA.

cs.IT

QoS Guaranteed Energy Minimization Over Two-Way Relaying Networks

In this work, we consider a typical three-node, two-way relaying network (TWRN) over fading channels. The aim is to minimize the entire system energy usage for a TWRN in the long run, while satisfying the required average symmetric exchange rate between the two source nodes. To this end, the energy usage of the physical-layer network coding (PNC) or the superposition coding based digital network coding (SPC-DNC) is analyzed. The rule on selection of both strategies is then derived by comparison. Based on the observed rule, we then design a scheme by switching between PNC and SPC-DNC for each channel realization. The associated optimization problem, through PNC/DNC switching,as well as power allocation on the uplink and the downlink for each channel realization is formulated and solved via an iterative algorithm. It is demonstrated that this switching scheme outperforms the schemes solely employing PNC or SPC-DNC through both theoretical analysis and simulations.

cs.IT

Joint Power Minimization Over Multi-Carrier Two-Way Relay Networks

The study considers a three-node, two-way relaying network (TWRN) over a multi-carrier system, aiming to minimize the total power consumption of all the transmit and receiver activities. By employing digital network coding (DNC) and physical-layer network coding (PNC), respectively, as well as a novel hybrid PNC/DNC switching scheme, the total transmission power of the considered multi-carrier TWRN system is firstly analyzed; and the derived analytical expressions are then used to formulate a set of nonconvex optimization problems. We will show how those nonconvex functions are convexified for better computational tracbility. It is observed in the numerical results that, the PNC scheme generally consumes less power than DNC but becomes worse in the very low SNR regime. The proposed hybrid PNC/DNC switching scheme that takes advantage of both, is shown to outperform in all SNR regimes.

cs.IT