arXiv ScienceSearch

arXiv subjects

Xinghang Zhang

Publications and source records attributed to Xinghang Zhang.

6 recordsLinked to original sources

Reducing Hallucinations in LLM-based Scientific Literature Analysis Using Peer Context Outlier Detection

Reducing hallucinations in Large Language Models (LLMs) is essential for accurate data extraction from large text corpora. Current methods, like prompt engineering and chain-of-thought prompting, focus on individual documents and fail to consider relationships across a corpus. This paper introduces Peer Context Outlier Detection (P-COD), which uses inter-document relationships to improve extraction accuracy in scientific literature summarization, where papers with similar experiment settings should draw similar conclusions. By comparing extracted data to validated peer information within the corpus, we adjust confidence scores and flag low-confidence results for expert review. Our experiments demonstrate up to 98% precision in outlier detection across 6 scientific domains, reducing hallucinations and letting researchers focus on genuinely ambiguous cases.

cs.AI

A Multi-Agent Human-LLM Collaborative Framework for Closed-Loop Scientific Literature Summarization

Scientific discovery is slowed by fragmented literature that requires excessive human effort to gather, analyze, and understand. AI tools, including autonomous summarization and question answering, have been developed to aid in understanding scientific literature. However, these tools lack the structured, multi-step approach necessary for extracting deep insights from scientific literature. Large Language Models (LLMs) offer new possibilities for literature analysis, but remain unreliable due to hallucinations and incomplete extraction. We introduce Elhuyar, a multi-agent, human-in-the-loop system that integrates LLMs, structured AI, and human scientists to extract, analyze, and iteratively refine insights from scientific literature. The framework distributes tasks among specialized agents for filtering papers, extracting data, fitting models, and summarizing findings, with human oversight ensuring reliability. The system generates structured reports with extracted data, visualizations, model equations, and text summaries, enabling deeper inquiry through iterative refinement. Deployed in materials science, it analyzed literature on tungsten under helium-ion irradiation, showing experimentally correlated exponential helium bubble growth with irradiation dose and temperature, offering insight for plasma-facing materials (PFMs) in fusion reactors. This demonstrates how AI-assisted literature review can uncover scientific patterns and accelerate discovery.

cs.AI

Lowering the Barrier to AI-Driven Inspection: A No-Code Workflow for Automated Structural Defect Detection

Structural health monitoring (SHM) is essential in modern engineering, providing data for condition-based maintenance, lifecycle assessment, and predictive decision-making. Traditionally, SHM relied on visual inspection to detect defects such as cracks and deformations. Early computer vision (CV) methods, including thresholding, edge detection, and handcrafted features, aimed to automate this process but were highly sensitive to noise, imaging variations, and multiscale defects, limiting their reliability. Recent advances in machine learning, particularly convolutional neural networks (CNNs) and You Only Look Once (YOLO), have improved defect detection accuracy and enabled real-time analysis. However, adoption in SHM remains limited due to technical barriers such as data labeling, model training, and deployment, which typically require programming expertise. To address this gap, we introduce YOLOEZ, an open-source, GUI-based tool for end-to-end YOLO model application. YOLOEZ integrates data labeling, training, and inference into a single interface, enabling high-performance model development without code while supporting reproducible workflows. Evaluation against existing software and classical image processing demonstrates that YOLOEZ not only outperforms traditional methods across most detection metrics, but also lowers adoption barriers present in other modern CV tools. By combining accuracy with accessibility, YOLOEZ facilitates wider use of AI-driven monitoring for predictive maintenance, digital twins, and intelligent structural systems.

cs.CV

Production of dileptons in ultra-peripheral heavy ion collisions with two-photon processes

We study the photoproduction process of dileptons in heavy ion collision at Relativistic Heavy Ion Collider (RHIC) and Large Hadron Collider (LHC) energys. The equivalent photon approximation, which equates the electromagnetic field of high-energy charged particles to the virtual photon flux, is used to calculate the processes of dileptons production. The numerical results demonstrate that the experimental study of dileptons in ultra-peripheral collisions is feasible at RHIC and LHC energies.

hep-ph

End-to-end Phase Field Model Discovery Combining Experimentation, Crowdsourcing, Simulation and Learning

The availability of tera-byte scale experiment data calls for AI driven approaches which automatically discover scientific models from data. Nonetheless, significant challenges present in AI-driven scientific discovery: (i) The annotation of large scale datasets requires fundamental re-thinking in developing scalable crowdsourcing tools. (ii) The learning of scientific models from data calls for innovations beyond black-box neural nets. (iii) Novel visualization and diagnosis tools are needed for the collaboration of experimental and theoretical physicists, and computer scientists. We present Phase-Field-Lab platform for end-to-end phase field model discovery, which automatically discovers phase field physics models from experiment data, integrating experimentation, crowdsourcing, simulation and learning. Phase-Field-Lab combines (i) a streamlined annotation tool which reduces the annotation time (by ~50-75%), while increasing annotation accuracy compared to baseline; (ii) an end-to-end neural model which automatically learns phase field models from data by embedding phase field simulation and existing domain knowledge into learning; and (iii) novel interfaces and visualizations to integrate our platform into the scientific discovery cycle of domain scientists. Our platform is deployed in the analysis of nano-structure evolution in materials under extreme conditions (high temperature and irradiation). Our approach reveals new properties of nano-void defects, which otherwise cannot be detected via manual analysis.

cs.CV

Cost-Effective Methods to Nanopattern Thermally Stable Platforms on Kapton HN Flexible Films Using Inkjet Printing Technology to Produce Printable Nitrate Sensors, Mercury Aptasensors, Protein Sensors, and Organic Thin Film Transistors

Kapton HN films, adopted worldwide due to their superior thermal durability (up to 400 °C), allow the high temperature sintering of nanoparticle based metal inks. By carefully selecting inks and Kapton substrates, outstanding thermal stability and anti-delaminating features are obtained in both aqueous and organic solutions and were applied to four novel devices: a solid state ion selective nitrate sensor, an ssDNA based mercury aptasensor, a low cost protein sensor, and a long lasting organic thin film transistor (OTFT). Many experimental studies on parameter combinations were conducted during the development of the above devices. The results showed that the ion selective nitrate sensor displayed a linear sensitivity range with a limit of detection of 2 ppm. The mercury sensor exhibited a linear correlation between the RCT values and the increasing concentrations of mercury. The protein printed circuit board (PCB) sensor provided a much simpler method of protein detection. Finally, the OTFT demonstrated a stable performance with mobility values for the linear and saturation regimes, and the threshold voltage. These devices have shown their value and reveal possibilities that could be pursued.

physics.app-ph