arXiv ScienceSearch

arXiv subjects

Jiatong Zhang

Publications and source records attributed to Jiatong Zhang.

8 recordsLinked to original sources

Caught in the Story: Narrative Captivity in Multi-turn LLMs Conversation

People increasingly turn to large language models (LLMs) for everyday advice, making ethically charged interpersonal problems a practical moral-advisory context. Most prior work has studied this context through single-turn judgments or pressure-laden rebuttals, assumptions that poorly match how guidance is sought in real-world contexts. These assumptions leave unclear whether narration alone, without an explicit opposing position, can shift model judgments during multi-turn moral consultation. Yet real-world moral-conflict conversation often elicits one party's self-justifying account, which can unfold over multiple turns and create information asymmetry. We introduce \textbf{narrative captivity}, a failure mode in which a model treats an unopposed one-sided account as complete and aligns with the narrator's interpretation without seeking missing perspectives. To measure this phenomenon, we build a benchmark of $5{,}078$ interpersonal-conflict scenarios spanning six moral dimensions. Across 17 LLMs, narrative captivity is widespread: end-state judgments under multi-turn narration shift by 25 percentage points on average beyond the matched single-turn baseline. Stage-level analysis identifies preference optimization as a major contributor, while four inference-time strategies provide only partial mitigation. We hope our project fosters LLM advisors that preserve independent judgment in real-world consultation.

cs.AI

Biodegradable, Millimeter-Scale Light-Emitting Sensors for Distributed Environmental Monitoring-Functional Pixie Dust

Methods for large-area, precise monitoring across natural environments are of growing interest due to pressing needs for sustainable management of rapidly increasing anthropogenic activities. Established approaches involve sparse spatial sampling and/or sequential measurements, while emerging techniques exploit miniaturized electronics or passive optical methods. Various constraints in scalability, costs, robustness, operational range and other factors create a need for alternatives. Here, we introduce a concept that overcomes many of these limitations through the combined use of chemically induced light emission and chemically responsive optical filter elements in millimeter-scale systems that we refer to as functional pixie dust (fPD) sensors, designed specifically for monitoring natural water systems during nighttime to eliminate background optical interference and to enhance remote analysis. These floating devices act as Lagrangian tracers to follow surface flows and to simultaneously measure the concentrations of key chemical species along their trajectories. Optimized designs exploit environmentally compatible constituent materials that are also degradable through natural processes to benign end products, thereby eliminating the need for recovery. Spatially and spectrally resolved ratiometric measurement schemes ensure robust operation and ability to address practical requirements in range, operational lifetime, time response and sensitivity. Demonstrations include distributed measurements of pH, Hg2+, and NO2-, each of relevance to industrial discharge, toxic metal contamination, and nitrogen-rich runoff, adapted for static concentration gradients, flow-driven transport conditions, and outdoor aquatic settings. The results establish a framework for environmental sensing using degradable, self-powered microsystems capable of scalable deployment and remote readout.

physics.app-ph

PRISM: Probing Reasoning, Instruction, and Source Memory in LLM Hallucinations

As large language models (LLMs) evolve from conversational assistants into agents capable of handling complex tasks, they are increasingly deployed in high-risk domains. However, existing benchmarks largely rely on mixed queries and posterior evaluation, output-level scoring, which quantifies hallucination severity but offers limited insight into where and why hallucinations arise in the generation pipeline. We therefore reformulate hallucination evaluation as a diagnostic problem and propose PRISM, a controlled benchmark that disentangles hallucinations into four dimensions: knowledge missing, knowledge errors, reasoning errors, and instruction-following errors, grounded in three stages of generation (memory, instruction, and reasoning). PRISM contains 9,448 instances across 65 tasks and supports fine-grained, stage-aware diagnostic evaluation. Evaluating 24 mainstream open-source and proprietary LLMs, we uncover consistent trade-offs across instruction following, memory retrieval, and logical reasoning, showing that mitigation strategies often improve specific dimensions at the expense of others. We hope PRISM provides a framework for understanding the specific mechanisms behind LLMs hallucinations, ultimately accelerating the development of trustworthy large language models.

cs.CL

Sandwich Reasoning: An Answer-Reasoning-Answer Approach for Low-Latency Query Correction

Query correction is a critical entry point in modern search pipelines, demanding high accuracy strictly within real-time latency constraints. Chain-of-Thought (CoT) reasoning improves accuracy but incurs prohibitive latency for real-time query correction. A potential solution is to output an answer before reasoning to reduce latency; however, under autoregressive decoding, the early answer is independent of subsequent reasoning, preventing the model from leveraging its reasoning capability to improve accuracy. To address this issue, we propose Sandwich Reasoning (SandwichR), a novel approach that explicitly aligns a fast initial answer with post-hoc reasoning, enabling low-latency query correction without sacrificing reasoning-aware accuracy. SandwichR follows an Answer-Reasoning-Answer paradigm, producing an initial correction, an explicit reasoning process, and a final refined correction. To align the initial answer with post-reasoning insights, we design a consistency-aware reinforcement learning (RL) strategy: a dedicated consistency reward enforces alignment between the initial and final corrections, while margin-based rejection sampling prioritizes borderline samples where reasoning drives the most impactful corrective gains. Additionally, we construct a high-quality query correction dataset, addressing the lack of specialized benchmarks for complex query correction. Experimental results demonstrate that SandwichR achieves SOTA accuracy comparable to standard CoT while delivering a 40-70% latency reduction, resolving the latency-accuracy trade-off in online search.

cs.AI

Magnetic properties of molecular beam epitaxy-grown ultrathin Cr2Ge2Te6 films down to monolayer limit on Si substrates

Cr2Ge2Te6, a prototypical van der Waals ferromagnetic semiconductor, have attracted significant interest for its potential applications in high-performance spintronics. However, the magnetic ground state of monolayer Cr2Ge2Te6 remains elusive due to fragile and irregular-shaped thin flake samples with weak magnetic signals. Here, we successfully grow uniform ferromagnetic Cr2Ge2Te6 films down to monolayer by molecular beam epitaxy. By exploiting a self-limiting growth mode, we achieve synthesis of uniform monolayer Cr2Ge2Te6 films across entire millimeter-scale Si substrates. Through a combination of superconducting quantum interference device magnetometry and anomalous Hall effect measurements, we establish that monolayer Cr2Ge2Te6 exhibits intrinsic ferromagnetism with perpendicular magnetic anisotropy below ~10 K, albeit with strong magnetic fluctuations characteristic of its two-dimensional nature. Furthermore, a systematic thickness-dependent study reveals a crossover from this fluctuation-dominated two-dimensional magnetism turns into conventional long-range ferromagnetic order as the film thickness increases. Our work not only definitively establishes the intrinsic ferromagnetic ground state of monolayer Cr2Ge2Te6, but also provides a scalable, silicon-compatible route for preparing the two-dimensional magnet for future spintronic or quantum devices.

cond-mat.mtrl-sci

Multimechanism quantum anomalous Hall and Chern number tunable states in germanene (silicene, stanene)/$M$Bi$_2$Te$_4$ heterostructures

By constructing germanene (silicene, stanene)/$M$Bi$_2$Te$_4$ ($M$ = 3d-transition elements) heterostructures, we discovered and designed multimechanism quantum-anomalous-Hall (QAH) systems, including $\Gamma$-based QAH, $K$-$K'$-connected QAH, and valley-polarized $K$- or $K'$-based QAH states via first-principle computations. The unique systems possess a global gap and tunable Chern number. The coexisting conventional $\Gamma$-based QAH state of $M$Bi$_2$Te$_4$ and valley-polarized $K$($K'$)-based QAH state of germanene (silicene, stanene), with opposite chirality, can interact with each other. Adjusting magnetic configurations of $M$Bi$_2$Te$_4$-layers not only switch on (off) the QAH conductance, but also modulate Chern numbers exactly. For example, the germanene/bilayer-NiBi$_2$Te$_4$ possesses the Chern number $C = +1$ in ferromagnetic couplings and $C = +2$ in antiferromagnetic couplings. The novel multimechanism QAH insulators, which are achievable in experiments, provide a new approach to spintronics and valleytronics based on topological states of matter.

cond-mat.mtrl-sci

Close-range Human Following Control on a Cane-type Robot with Multi-camera Fusion

Cane-type robots have been utilized to assist and supervise the mobility-impaired population. One essential technique for cane-type robots is human following control, which allows the robot to follow the user. However, the limited perceptible information of humans by sensors at close range, combined with the occlusion caused by lower limb swing during normal walking, affect the localization of users. These limitations make it difficult to achieve human following at close range.To address these challenges, this study developed a new cane-type wheeled robot and proposed a novel human-following control with multi-camera fusion. This control system mainly consists of two parts: 1) a human following controller that locates a user by multi-camera fusion and generates control signals to follow the user. 2) a cane robot controller designed to steer the cane robot to a target position. The proposed strategy's effectiveness has been validated in outdoor experiments with six healthy subjects. The experimental scenarios included different terrains (i.e., straight, turning, and inclined paths), road conditions (i.e., flat and rough roads), and walking speeds. The obtained results showed that the average tracking error for position and orientation was less than 5 cm and 15{\deg} respectively across all scenarios. Moreover, the cane robot can effectively adapt to a wide range of individual gait patterns and achieve stable human following at daily walking speeds (0.75 m/s - 1.45 m/s).

eess.SY

Semantic-Aware Local-Global Vision Transformer

Vision Transformers have achieved remarkable progresses, among which Swin Transformer has demonstrated the tremendous potential of Transformer for vision tasks. It surmounts the key challenge of high computational complexity by performing local self-attention within shifted windows. In this work we propose the Semantic-Aware Local-Global Vision Transformer (SALG), to further investigate two potential improvements towards Swin Transformer. First, unlike Swin Transformer that performs uniform partition to produce equal size of regular windows for local self-attention, our SALG performs semantic segmentation in an unsupervised way to explore the underlying semantic priors in the image. As a result, each segmented region can correspond to a semantically meaningful part in the image, potentially leading to more effective features within each of segmented regions. Second, instead of only performing local self-attention within local windows as Swin Transformer does, the proposed SALG performs both 1) local intra-region self-attention for learning fine-grained features within each region and 2) global inter-region feature propagation for modeling global dependencies among all regions. Consequently, our model is able to obtain the global view when learning features for each token, which is the essential advantage of Transformer. Owing to the explicit modeling of the semantic priors and the proposed local-global modeling mechanism, our SALG is particularly advantageous for small-scale models when the modeling capacity is not sufficient for other models to learn semantics implicitly. Extensive experiments across various vision tasks demonstrates the merit of our model over other vision Transformers, especially in the small-scale modeling scenarios.

cs.CV