arXiv ScienceSearch

arXiv subjects

Jungwoo Lee

Publications and source records attributed to Jungwoo Lee.

At least 19 recordsLinked to original sources

Refining transposition bounds from traceless quantum perturbations

The transpose map is the canonical nonphysical map in quantum information that is central to entanglement tests and to converse bounds for quantum capacity. It is a longstanding conjecture that transposition satisfies a dimension-independent strict submultiplicativity for physically relevant channel differences. We answer this affirmatively by proving that for every quantum channel $T$ on $\mathsf{L}(\mathbb{C}^d)$, $\|Θ\circ(\mathrm{id}-T)\|_\diamond \le \frac{1}{\sqrt{2}}\|Θ\|_\diamond \|\mathrm{id}-T\|_\diamond$. The same bound holds for all Hermiticity-preserving trace-annihilating maps. Because a trace-preserving channel produces a traceless deviation from the identity, transposition cannot act as strongly on such physically generated errors as it can on arbitrary perturbations. Thus, although transposition is extremal in the unrestricted diamond-norm geometry, physical constraints mandate a universal margin in the submultiplicative bound. As one immediate consequence, the finite-error Holevo--Werner converse extends from decoding errors $\varepsilon<1/2$ to $\varepsilon<1/\sqrt{2}$. The same tracelessness mechanism also yields rank-sensitive continuity refinements for negativity and logarithmic negativity, demonstrating its potential for broader application in quantum information.

quant-ph

One-parameter counterexamples to the refined Bessis-Moussa-Villani conjecture

Positivity of matrix trace exponentials is a basic structural principle behind finite-temperature quantum statistical mechanics. The Bessis-Moussa-Villani conjecture, a central manifestation of this principle, was proved by Stahl after an influential reformulation by Lieb and Seiringer. A later refinement asks whether the normalized average over all words with $n$ letters $A$ and $m$ letters $B$ is always bounded above by $\mathrm{tr}(A^nB^m)$ and below by $\mathrm{tr}\exp(n\log A+m\log B)$. In this work, we study a specific one-parameter family $(A_x, B_x)$ and show that the correct small-$x$ invariant of a word is not its degree of fragmentation, but a weighted shortest-bridge cost on its cyclic run decomposition. Our results yield a class of counterexamples to the suggested refinement. Remarkably, the ratio of the normalized word average to the trace $\mathrm{tr}(A^nB^m)$ can become arbitrarily large.

quant-ph

Towards Understanding Pause Token Fine-Tuning Dynamics: A Mode Retention Perspective

Pause-token methods improve LLM reasoning by inserting special tokens into sequences. Prior work explains these gains through computational expressivity. However, there is relatively little investigation into the training dynamics of pause tokens. We explore how pause tokens reshape the training dynamics of fine-tuning. Two controlled pilots expose distinct asymmetries. On a synthetic continual-learning task, masked pauses overwrite a previously-learned distribution roughly 4x less at matched final adaptation (H1, mode retention); on a synthetic math-reasoning probe, the boundary-adjacent token comes to encode substantially more downstream-step information (H2, non-myopic compression). We formalize a training rule consistent with both - Masked Boundary Pause (MBP), pause tokens placed at reasoning-step boundaries with their loss masked. Across 1B-8B Qwen and Llama models, MBP consistently improves reasoning, achieving gains of up to 6 points on math and 2.5 points on code, while preserving general language understanding abilities. We further demonstrate that this mode-preserving strategy extend gains to GRPO. These results recast pause tokens as a training-dynamics intervention on the retention-adaptation trade-off, rather than merely an inference-time computation device.

cs.CL

Online Estimation of Partial Transpose Moments via Fast Classical Updates

Partial-transpose (PT) moments are among the most practically relevant nonlinear quantities accessible from local Pauli classical shadows, because they directly underpin mixed-state entanglement certification and recent PT-moment-based phase diagnostics. The online framework of Marso \emph{et al.} rewrote the exact PT-moment statistic into a fixed-memory recurrence that updates a small collection of accumulated matrices after each new shadow snapshot. Its update cost is independent of the shot number, but each step treats the incoming partially transposed snapshot as a generic dense matrix. Therefore, the arithmetic cost scales cubically with the dimension of the Hilbert space. We show that the same estimator can be updated exactly in subcubic time per shot while retaining the same memory. The key point is that the accumulated matrices become dense, but the fresh partially transposed snapshot still factorizes into local factors. Right-multiplication by that factorized snapshot can therefore be executed by exact column-pair sweeps. For the second PT moment, we further optimize the process by utilizing a Pauli basis update.

quant-ph

Exactness of the doubly nonnegative relaxation for qubit-output quantum channels

The resource theory for nonnegativity of quantum amplitudes distinguishes completely positive completely positive (CPCP) quantum channels from the larger class of completely positive doubly nonnegative (CPDNN) quantum channels. Johnston and Sikora showed that all qubit-to-qubit quantum channels that are CPDNN are also CPCP. However, they left open the question of whether a qutrit-to-qubit quantum channel exists that is CPDNN but not CPCP. We prove that no such channel exists and, more generally, that every CPDNN quantum channel with qubit output is CPCP. Thus the doubly nonnegative relaxation is exact for all qubit-output quantum channels. Our argument yields an explicit structural picture in which, after a canonical permutation of the Choi matrix, every qubit-output CPDNN channel is determined by a nonnegative vector and a nonnegative matrix, subject to a single positive-semidefinite block constraint. This gives a complete binary-output normal form, an explicit formula for the action of the channel, and quantitative population-coherence tradeoff inequalities governing the unique output off-diagonal mode. In this sense, qubit output is the nontrivial binary regime in which trace preservation and double nonnegativity collapse the free-channel cone to a fully describable geometry.

quant-ph

Disjoint Bell measurements enable near-projective GHZ certification

Certifying multipartite entangled states is a basic task in quantum information processing, but the achievable copy complexity depends crucially on the measurements available to the verifier. The strongest possible certification measurement for a known pure target state $|ψ\rangle$ is the two-outcome projector $\{|ψ\rangle\langleψ|,\mathsf{I}-|ψ\rangle\langleψ|\}$, which is copy-optimal but often experimentally unrealistic or outside the intended measurement model. In this work, we introduce BM-Cert, a single-copy verification protocol for the $n$-qubit Greenberger--Horne--Zeilinger (GHZ) state using only disjoint two-qubit Bell-basis measurements, together with one single-qubit $X$-basis measurement when $n$ is odd. Surprisingly, a simple combinatorial effect yields perfect completeness and a verification spectral gap $ν_\mathrm{BM}(n)=1-O(1/n)$, so our depth-2 protocol already approaches the ideal projective verification asymptotically as $n$ grows. This contrasts with local Pauli GHZ verification, whose optimal spectral gap remains bounded away from $1$. Thus, allowing only two-qubit entangling measurements on disjoint pairs is enough to achieve asymptotically ideal projective certification. The same Bell-matching outcomes also yield BM-Fid, an unbiased estimator of the GHZ fidelity whose leading Hoeffding coefficient in the sample complexity tends to the ideal value achieved by direct projection. For the open-boundary linear nearest neighbor setting, we further introduce Brick-Cert, a disjoint 2-local certification protocol whose spectral gap $4/5$ is optimal within that restricted architecture.

quant-ph

Every architecture of six two-qubit gates is locally universal on three qubits

We present the \emph{first} analytical determination of the exact accessible dimension for every fixed architecture of arbitrary two-qubit gates on three qubits. A support word represents which pair of qubits each two-qubit gate acts on, and its reduced length is obtained by merging consecutive gates on the same pair. If the reduced length is $r$, then the set of implementable three-qubit unitaries has accessible dimension $d(w)=\min\{63,9r+9\}$. Consequently, six arbitrary two-qubit gates are \emph{necessary and sufficient for local universality}: every fixed architecture reaches a nonempty open subset of $\mathrm{SU}(8)$ when its reduced length is at least six. The result is stronger than existence of one favorable architecture. Every reduced six-slot support word is locally universal, including the alternating nearest-neighbor line $AB,BC,AB,BC,AB,BC$. The upper bound follows from a standard parameter-counting argument. The matching lower bounds are proved by Jacobian certificates in the Pauli basis. Up to qubit relabeling and reversal, there are $22$ reduced architectures of lengths two through six. For each one, a product of rational Pauli rotations yields a nonzero maximal minor modulo the prime $1{,}000{,}003$. Thus, any obstruction to global universality with six two-qubit gates must go beyond parameter counting, connectivity, and differential rank.

quant-ph

DCGC: Draft-Conditioned Global Correction for Complex Reasoning with Masked Diffusion Models

Correcting flawed reasoning traces remains a significant challenge for Large Language Models (LLMs), whose autoregressive generation can propagate early mistakes into subsequent reasoning. We introduce DCGC, a Masked Diffusion Model (MDM) framework for global correction that uses an imperfect solution draft from an upstream solver as auxiliary context. DCGC combines task-specific Supervised Fine-Tuning (SFT) with a novel inference-time mechanism called Dynamic Dual-CFG. This mechanism separates problem-only and joint problem-draft branches and scales the draft-conditioned residual using a relative confidence gap. Across math, code, and knowledge reasoning benchmarks, DCGC outperforms standard sampling and simpler CFG variants, with additional results suggesting transfer to different diffusion backbones. In test-time setting where ground-truth failure labels are unavailable, DCGC improves full test set accuracy by correcting low-consensus upstream outputs, highlighting its utility as a verifier-free global correction module for difficult reasoning instances.

cs.CL

Mechanistic Circuit Identification for Controllable Data Generation

While recent advances in data synthesis aim to curate high-quality datasets, most generation pipelines still rely on heuristic prompt-based control. This black-box paradigm provides limited insight into how individual samples interact with a model's underlying learning dynamics. To bridge this gap, we propose a circuit-grounded framework that connects training-dynamics-based data valuation with mechanistic interpretability (MI). Specifically, we conceptualize data quality along three complementary utility axes, learnability, challenge, and alignment. First, we uncover specialized model-internal circuits that causally govern these utility signals. Then, moving beyond heuristic prompting toward mechanistic control, we leverage these circuits as controllable interfaces, actively steering generation to produce utility-targeted data. Building on this capability, we introduce SAMS (Stage-Aware Mechanistic Scheduling), which schedules circuit-steered data according to the model's evolving optimization needs. Experiments on multiple-choice QA tasks demonstrate that our approach yields precisely controlled data with greater diversity than prompt-based baselines, consistently improving downstream performance and calibration. Ultimately, this work establishes a principled white-box paradigm for interpretable data generation, pioneering the use of MI not just as an analytical tool, but as a practical, controllable interface.

cs.LG

Structural perspectives from quantum states and measurements in optimal state discrimination

Quantum state discrimination refers to a class of techniques to identify a specific quantum state through a \textit{positive operator-valued measure}. In this work, we investigate how structural information can influence our ability to determine or bound the optimal discrimination probability. First, as background, we note that for single-qubit pure-state ensembles, pairwise fidelities determine the optimal discrimination probability, whereas in higher dimensions they do not in general. As an illustrative application of this observation, we give a closed-form fidelity-based reformulation of the optimal discrimination probability for three equiprobable single-qubit states with equal pairwise fidelities. Secondly, we show that the information of measurement operators that vanish in the optimal solution can be used to refine upper bounds on the optimal discrimination probability, often yielding tighter bounds.

quant-ph

Function-Level Execution Feedback for Code Preference Optimization

Process supervision has improved mathematical reasoning, where intermediate steps are naturally expressed as chains of thought. In code generation, however, process supervision remains underexplored because there is no standard notion of a step. Supervision can target lines, reasoning traces, or program states, making it unclear what to label and optimize. We propose STEP-KTODER, a framework for code preference optimization that defines steps as module-level functions in decomposed multi-function programs and assigns binary correctness labels via automatically generated unit tests. Our method provides a code-specific instantiation of stepwise KTO, combining function-level process supervision with outcome-level feedback on the full program. We evaluate on HumanEval(+), MBPP(+), BigCodeBench, and LiveCodeBench, showing that STEP-KTODER improves over outcome-only KTO and DPO. Further analysis shows that execution-based labels are essential: LLM-as-a-judge annotations systematically over-predict function failures, corrupt positive step labels, and degrade downstream preference optimization. Code is available at: https://github.com/inechnech/STEP-KTODER.

cs.AI

Re-uploading quantum data: a universal function approximator for quantum inputs

Quantum data re-uploading has proved powerful for classical inputs, where repeatedly encoding features into a small circuit yields universal function approximation. Extending this idea to quantum inputs remains underexplored, as the information contained in a quantum state is not directly accessible in classical form. We propose and analyze a quantum data re-uploading architecture in which a qubit interacts sequentially with fresh copies of an arbitrary input state. The circuit can approximate any bounded continuous function using only one ancilla qubit and single-qubit measurements. By alternating entangling unitaries with mid-circuit resets of the input register, the architecture realizes a discrete cascade of completely positive and trace-preserving maps, analogous to collision models in open quantum system dynamics. Our framework provides a qubit-efficient and expressive approach to designing quantum machine learning models that operate directly on quantum data.

quant-ph

MSG-Loc: Multi-Label Likelihood-based Semantic Graph Matching for Object-Level Global Localization

Robots are often required to localize in environments with unknown object classes and semantic ambiguity. However, when performing global localization using semantic objects, high semantic ambiguity intensifies object misclassification and increases the likelihood of incorrect associations, which in turn can cause significant errors in the estimated pose. Thus, in this letter, we propose a multi-label likelihood-based semantic graph matching framework for object-level global localization. The key idea is to exploit multi-label graph representations, rather than single-label alternatives, to capture and leverage the inherent semantic context of object observations. Based on these representations, our approach enhances semantic correspondence across graphs by combining the likelihood of each node with the maximum likelihood of its neighbors via context-aware likelihood propagation. For rigorous validation, data association and pose estimation performance are evaluated under both closed-set and open-set detection configurations. In addition, we demonstrate the scalability of our approach to large-vocabulary object categories in both real-world indoor scenes and synthetic environments. Project Page: https://sparolab.github.io/research/msg-loc/.

cs.RO

RadLoc: Radar-based 3-DoF Global Localization via Fast, Robust, and Lightweight Spatial Descriptor Across Diverse Environmental Scenarios

While global localization using spinning radar has gained attention for its robustness to adverse weather and challenging environments, many studies have focused on individual components such as place recognition or pose estimation. In this paper, we take a holistic view of radar sensor-based global localization and present RadLoc, a fast, robust, and lightweight end-to-end pipeline from place recognition to 3-DoF pose estimation. RadLoc accelerates pre-processing using 1D CA-CFAR filtering and leverages the near-range dominance in spinning radar images to design a compact descriptor and an efficient hierarchical coarse-to-fine retrieval strategy. Moreover, coupled with phase correlation-based 3-DoF pose estimation, it forms a versatile global localization module applicable to SLAM and multi-session SLAM systems. Extensive experiments on 15 sequences across 5 datasets demonstrate that RadLoc achieves robust performance while maintaining the smallest descriptor size and fastest retrieval time among state-of-the-art approaches. The supplementary materials are available at https://sparolab.github.io/research/radloc/.

cs.RO

MixTTA: Low-Rank Cross-Channel Mixing for Reliable Test-Time Adaptation

Test-Time Adaptation (TTA) methods commonly update the affine parameters of normalization layers to adapt deployed models under distribution shifts. However, per-channel affine parameters perform axis-aligned scaling and shifting, making them geometrically incapable of correcting cross-channel structural changes induced by distribution shift. To address this limitation, we propose MixTTA, a lightweight plug-in module that equips normalization layers with a low-rank cross-channel transformation, enabling inter-channel mixing at each layer. To ensure that the low-rank branch captures only cross-channel interactions, we also propose Decoupling Projection that enforces strict separation from the diagonal affine path, along with Spectral Projection that prevents rank-1 collapse under non-stationary test streams. MixTTA can be seamlessly integrated into any existing normalization-based TTA method. Experiments in both standard and wild TTA settings show consistent improvements over strong baselines while mitigating adaptation failure under challenging conditions. The source code is publicly available at https://github.com/delta6189/MixTTA.

cs.LG

Optimal dense materialization of the stabilizer formalism without polynomial overhead

Stabilizer states and Clifford transformations constitute the exactly tractable backbone of quantum information science, from error correction and fault tolerance to benchmarking and simulation. Although these objects admit compact classical descriptions, many physical and computational workflows still require their explicit dense forms such as a full wavefunction for a stabilizer state or a full matrix for a Clifford transformation. In such explicit output tasks, exponential scaling is unavoidable because the outputs themselves have sizes $2^n$ and $4^n$. The fundamental question is therefore whether compact stabilizer and Clifford descriptions can be expanded with no additional polynomial overhead. Here we answer this question affirmatively. We present optimal algorithms that materialize an $n$-qubit stabilizer state vector in $O(2^n)$ time and a full dense Clifford matrix in $O(4^n)$ time. The same framework also yields an optimal conversion from standard stabilizer check matrices to state vectors and, for every fixed odd prime qudit dimension $\ell$, gives $O(\ell^n)$-time materialization of qudit stabilizer states. As an additional compact-to-compact result, we design a sign-aware Four Russians method for converting stabilizer check matrices to quadratic forms faster than Gaussian elimination. These results close the asymptotic gap between compact descriptions of the stabilizer formalism and their dense representations.

quant-ph

A Regret Minimization Framework on Preference Learning in Large Language Models

Reinforcement learning with verifiable rewards (RLVR) has enabled progress on reasoning-intensive tasks by relying on task-specific verifiers that provide automated correctness signals. However, many realistic language tasks are difficult to equip with reliable verifiers, motivating a growing reliance on reinforcement learning from human feedback (RLHF). In this setting, we argue that a closer examination of how human feedback should be interpreted is essential. We introduce Regret-based Preference Optimization $(\textbf{RePO})$, which reframes RLHF through $\textit{regret minimization}$ rather than reward maximization. Human preferences are often shaped by $\textit{prospective}$ anticipation of outcomes and $\textit{counterfactual}$ comparisons to alternative behaviors, rather than by immediate, outcome-independent utility. $\textbf{RePO}$ captures this structure by modeling preferences as behavior-conditioned assessments of relative suboptimality. Experiments on mathematical reasoning benchmarks and human preference datasets demonstrate consistent performance gains, indicating that $\textbf{RePO}$ is an effective and human-aligned approach for training large language models.

cs.AI

MVP-LAM: Learning Action-Centric Latent Action via Cross-Viewpoint Reconstruction

Latent actions learned from diverse human videos serve as pseudo-labels for vision-language-action (VLA) pretraining, but provide effective supervision only if they remain informative about the underlying ground-truth actions. For effective supervision, latent actions should contain information about the underlying actions even though they are inaccessible. We propose Multi-ViewPoint Latent Action Moel (MVP-LAM), which learns latent actions that are highly informative about ground-truth actions from multi-view videos. MVP-LAM trains latent actions with a cross-viewpoint reconstruction objective, so that a latent action from one view must explain the future in another view, reducing reliance on viewpoint-specific cues. On Bridge V2, MVP-LAM produces more action-centric latent actions, achieving higher mutual information with ground-truth actions and improved action prediction, including under out-of-distribution evaluation. Finally, pretraining VLAs with MVP-LAM latent actions improves downstream manipulation performance on various benchmarks. The code and trained checkpoints are available at https://jmsnu.github.io.

cs.RO