arXiv ScienceSearch

arXiv subjects

Zihe Wang

Publications and source records attributed to Zihe Wang.

At least 19 recordsLinked to original sources

Random Forest-Based Prediction of Bone Volume Fraction and Fracture Position from S-Parameters

In this paper, we propose a method for predicting bone volume fraction (BVF) and fracture position by constructing a random forest model based on multichannel S-parameters. A nine-antenna microwave scanning system is designed and fabricated to acquire the multichannel S-parameter data. Bone-mimicking phantoms are developed, and corresponding experiments are conducted to validate the effectiveness of the proposed approach. Both synthetic and experimental results demonstrate the validity of the method.

cs.LG

Private Private Information in Second-Price Auction

Classic results show that even an arbitrarily small correlation across bidders' information can enable full surplus extraction in auctions and related mechanism design settings. Motivated by this fragility, we study the information independence in a second-price auction when the seller commits to a private private information structure, meaning bidders' signals are independent ex ante, while bidders share a symmetric and arbitrarily correlated prior distribution over their valuations. We first show that the seller optimal efficient outcome with full surplus extraction can always be implemented by a private private information structure that admits a Bayes Nash equilibrium. However, this equilibrium may not be stable. We then further construct a private private information structure that achieves revenue arbitrarily close to maximum welfare while admitting a strict equilibrium. At the same time, we establish an impossibility result: under private private information, in general, bidder surplus cannot achieve maximal welfare exactly, and we characterize necessary and sufficient conditions on the prior distribution under which bidder surplus can be made arbitrarily close to maximal welfare. We finally explore which other efficient outcomes are achievable under private private information.

econ.TH

Algorithmic Information Design for Searchers with Uncertain Alternatives

Advertisements reveal information to consumers who decide on further information acquisition and eventual purchase. Anderson and Renault (2006) first modeled this problem using an information-design framework where the advertiser acts as a sender and the consumer as a receiver. Due to search frictions and the consumer's outside option, search for additional information is not always worthwhile for the receiver. Thus, the sender's information design is used to make further search attractive and ultimately induce purchase. Following Lyu (2023), we study the optimal information design for a sender who sells search goods to a consumer with uncertain alternatives. Instead of making relaxations to the sender's problem (as done in Lyu, 2023), we work directly on the joint distribution over realized values and signals. Our contributions are twofold. First, we give a method, based on duality arguments, to verify whether a given information strategy is optimal. We illustrate the value of this verification framework in a competitive extension, where it certifies a non-trivial symmetric equilibrium for two senders with a common convex prior. Second, on the algorithmic front, we develop an FPTAS that finds for the seller a signaling scheme whose utility differs from that of the optimal solution by at most an additive $ε$ error, for any $ε> 0$.

cs.GT

Second-Best Bilateral Trade is $1/2$ Efficient

The landmark Myerson-Satterthwaite Theorem establishes a fundamental impossibility in bilateral trade: no Bayesian incentive-compatible mechanism can simultaneously achieve ex-post efficiency, individual rationality, and strong budget balance. We resolve a long-standing open question regarding the efficiency loss imposed by these constraints. Specifically, we prove that the Bayesian-optimal (second-best) mechanism always captures at least half of the first-best gains from trade ($\mathrm{SB}\ge\frac{1}{2}\mathrm{FB}$). This result is tight, definitively closing the gap between the previously best-known bounds of $0.317$ and $0.736$.

cs.GT

Pacing Equilibria in Second-Price Auctions with Few Goods

In this paper, we investigate the computation of second-price pacing equilibria (SPPEs), a foundational model in online advertising auctions. We present a polynomial-time algorithm for computing exact SPPEs in instances with a constant number of goods. Our core technique maps buyers' pacing multipliers to the highest bids on each good, effectively partitioning the parameter space into a set of distinct geometric cells. By enumerating these cells, we fix the relative ordering of the bids and reduce the problem of equilibrium computation to a linear feasibility program. Finally, we demonstrate that this tractability extends to large-scale markets with an arbitrary number of goods, provided the goods can be aggregated into a constant number of valuation types.

cs.GT

AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models

The rise of vision foundation models (VFMs) calls for systematic evaluation. A common approach pairs VFMs with large language models (LLMs) as general-purpose heads, followed by evaluation on broad Visual Question Answering (VQA) benchmarks. However, this protocol has two key blind spots: (i) the instruction tuning data may not align with VQA test distributions, meaning a wrong prediction can stem from such data mismatch rather than a VFM' visual shortcomings; (ii) VQA benchmarks often require multiple visual abilities, making it hard to tell whether errors stem from lacking all required abilities or just a single critical one. To address these gaps, we introduce AVA-Bench, the first benchmark that explicitly disentangles 14 Atomic Visual Abilities (AVAs) -- foundational skills like localization, depth estimation, and spatial understanding that collectively support complex visual reasoning tasks. By decoupling AVAs and matching training and test distributions within each, AVA-Bench pinpoints exactly where a VFM excels or falters. Applying AVA-Bench to leading VFMs thus reveals distinctive "ability fingerprints," turning VFM selection from educated guesswork into principled engineering. Notably, we find that a 0.5B LLM yields similar VFM rankings as a 7B LLM while cutting GPU hours by 8x, enabling more efficient evaluation. By offering a comprehensive and transparent benchmark, we hope AVA-Bench lays the foundation for the next generation of VFMs.

cs.CV

ExpressMind: A Multimodal Pretrained Large Language Model for Expressway Operation

The current expressway operation relies on rule-based and isolated models, which limits the ability to jointly analyze knowledge across different systems. Meanwhile, Large Language Models (LLMs) are increasingly applied in intelligent transportation, advancing traffic models from algorithmic to cognitive intelligence. However, general LLMs are unable to effectively understand the regulations and causal relationships of events in unconventional scenarios in the expressway field. Therefore, this paper constructs a pre-trained multimodal large language model (MLLM) for expressways, ExpressMind, which serves as the cognitive core for intelligent expressway operations. This paper constructs the industry's first full-stack expressway dataset, encompassing traffic knowledge texts, emergency reasoning chains, and annotated video events to overcome data scarcity. This paper proposes a dual-layer LLM pre-training paradigm based on self-supervised training and unsupervised learning. Additionally, this study introduces a Graph-Augmented RAG framework to dynamically index the expressway knowledge base. To enhance reasoning for expressway incident response strategies, we develop a RL-aligned Chain-of-Thought (RL-CoT) mechanism that enforces consistency between model reasoning and expert problem-solving heuristics for incident handling. Finally, ExpressMind integrates a cross-modal encoder to align the dynamic feature sequences under the visual and textual channels, enabling it to understand traffic scenes in both video and image modalities. Extensive experiments on our newly released multi-modal expressway benchmark demonstrate that ExpressMind comprehensively outperforms existing baselines in event detection, safety response generation, and complex traffic analysis. The code and data are available at: https://wanderhee.github.io/ExpressMind/.

cs.AI

Mechanism Design via Market Clearing-Prices for Value Maximizers under Budget and RoS Constraints

The transition to auto-bidding in online advertising has shifted the focus of auction theory from quasi-linear utility maximization to value maximization subject to financial constraints. We study mechanism design for buyers with private budgets and private Return-on-Spend (RoS) constraints, but public valuations, a setting motivated by modern advertising platforms where valuations are predicted via machine learning models. We introduce the extended Eisenberg-Gale program, a convex optimization framework generalized to incorporate RoS constraints. We demonstrate that the solution to this program is unique and characterizes the market's competitive equilibrium. Based on this theoretical analysis, we design a market-clearing mechanism and prove two key properties: (1) it is incentive-compatible with respect to financial constraints, making truthful reporting the optimal strategy; and (2) it achieves a tight 1/2-approximation of the first-best revenue benchmark, the maximum revenue of any feasible mechanism, regardless of IC. Finally, to enable practical implementation, we present a decentralized online algorithm. Ignoring logarithmic factors, we prove that under this algorithm, both the seller's revenue and each buyer's utility converge to the equilibrium benchmarks with a sublinear regret of $\tilde{O}(\sqrt{m})$ over $m$ auctions.

cs.GT

Intelli-Planner: Towards Customized Urban Planning via Large Language Model Empowered Reinforcement Learning

Effective urban planning is crucial for enhancing residents' quality of life and ensuring societal stability, playing a pivotal role in the sustainable development of cities. Current planning methods heavily rely on human experts, which are time-consuming and labor-intensive, or utilize deep learning algorithms, often limiting stakeholder involvement. To bridge these gaps, we propose Intelli-Planner, a novel framework integrating Deep Reinforcement Learning (DRL) with large language models (LLMs) to facilitate participatory and customized planning scheme generation. Intelli-Planner utilizes demographic, geographic data, and planning preferences to determine high-level planning requirements and demands for each functional type. During training, a knowledge enhancement module is employed to enhance the decision-making capability of the policy network. Additionally, we establish a multi-dimensional evaluation system and employ LLM-based stakeholders for satisfaction scoring. Experimental validation across diverse urban settings shows that Intelli-Planner surpasses traditional baselines and achieves comparable performance to state-of-the-art DRL-based methods in objective metrics, while enhancing stakeholder satisfaction and convergence speed. These findings underscore the effectiveness and superiority of our framework, highlighting the potential for integrating the latest advancements in LLMs with DRL approaches to revolutionize tasks related to functional areas planning.

cs.AI

Deterministic implementation in single-item auctions

Deterministic auctions are attractive in practice due to their transparency, simplicity, and ease of implementation, motivating a sharper understanding of when they can attain the same outcomes as randomized mechanisms. We study deterministic implementation in single-item auctions under two notions of outcomes: (revenue, welfare) pairs and interim allocations. For (revenue, welfare) pairs, we show a separation in discrete settings: there exists a pair implementable by a deterministic Bayesian incentive-compatible (BIC) auction but not by any deterministic dominant-strategy incentive-compatible (DSIC) auction. For continuous atomless priors, we identify conditions under which deterministic DSIC auctions are equivalent to randomized BIC auctions in terms of achievable outcomes. For interim allocations, under a strict monotonicity condition, we establish a deterministic analogue of Border's theorem for two bidders, providing a necessary and sufficient condition for deterministic DSIC implementability. Using this characterization, we exhibit an interim allocation implementable by a randomized BIC auction but not by any deterministic DSIC auction.

cs.GT

Are High-Degree Representations Really Unnecessary in Equivariant Graph Neural Networks?

Equivariant Graph Neural Networks (GNNs) that incorporate E(3) symmetry have achieved significant success in various scientific applications. As one of the most successful models, EGNN leverages a simple scalarization technique to perform equivariant message passing over only Cartesian vectors (i.e., 1st-degree steerable vectors), enjoying greater efficiency and efficacy compared to equivariant GNNs using higher-degree steerable vectors. This success suggests that higher-degree representations might be unnecessary. In this paper, we disprove this hypothesis by exploring the expressivity of equivariant GNNs on symmetric structures, including $k$-fold rotations and regular polyhedra. We theoretically demonstrate that equivariant GNNs will always degenerate to a zero function if the degree of the output representations is fixed to 1 or other specific values. Based on this theoretical insight, we propose HEGNN, a high-degree version of EGNN to increase the expressivity by incorporating high-degree steerable vectors while maintaining EGNN's efficiency through the scalarization trick. Our extensive experiments demonstrate that HEGNN not only aligns with our theoretical analyses on toy datasets consisting of symmetric structures, but also shows substantial improvements on more complicated datasets such as $N$-body and MD17. Our theoretical findings and empirical results potentially open up new possibilities for the research of equivariant GNNs.

cs.LG

Universally Invariant Learning in Equivariant GNNs

Equivariant Graph Neural Networks (GNNs) have demonstrated significant success across various applications. To achieve completeness -- that is, the universal approximation property over the space of equivariant functions -- the network must effectively capture the intricate multi-body interactions among different nodes. Prior methods attain this via deeper architectures, augmented body orders, or increased degrees of steerable features, often at high computational cost and without polynomial-time solutions. In this work, we present a theoretically grounded framework for constructing complete equivariant GNNs that is both efficient and practical. We prove that a complete equivariant GNN can be achieved through two key components: 1) a complete scalar function, referred to as the canonical form of the geometric graph; and 2) a full-rank steerable basis set. Leveraging this finding, we propose an efficient algorithm for constructing complete equivariant GNNs based on two common models: EGNN and TFN. Empirical results demonstrate that our model demonstrates superior completeness and excellent performance with only a few layers, thereby significantly reducing computational overhead while maintaining strong practical efficacy.

cs.LG

Multiplayer General Lotto game

In this paper, we investigate the multiplayer General Lotto game across multiple battlefields, a significant variant of the Colonel Blotto game. In this version, each player employs a probability distribution for resource allocation, ensuring that their expected expenditure does not exceed their budget. We first establish the existence of the Nash equilibrium in a general setting, where players' budgets are asymmetric and the values of the battlefields are heterogeneous and asymmetric among players. Next, we provide a detailed characterization of the Nash equilibrium for multiple players on a single battlefield. In this characterization, we observe that the upper endpoints of the supports of players' equilibrium strategies coincide, and that the minimum value of a player's support above zero inversely correlates with his budget. We demonstrate the uniqueness of Nash equilibrium over a single battlefield in some scenarios. In the multi-battlefield setting, we prove that there is an upper bound on the average number of battlefields each player participates in. Additionally, we provide an example demonstrating the non-uniqueness of the Nash equilibrium in the context of multiple battlefields with multiple players. Finally, we present a solution for the Nash equilibrium in a symmetric case.

cs.GT

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs

Game-playing ability serves as an indicator for evaluating the strategic reasoning capability of large language models (LLMs). While most existing studies rely on utility performance metrics, which are not robust enough due to variations in opponent behavior and game structure. To address this limitation, we propose \textbf{Cognitive Hierarchy Benchmark (CHBench)}, a novel evaluation framework inspired by the cognitive hierarchy models from behavioral economics. We hypothesize that agents have bounded rationality -- different agents behave at varying reasoning depths/levels. We evaluate LLMs' strategic reasoning through a three-phase systematic framework, utilizing behavioral data from six state-of-the-art LLMs across fifteen carefully selected normal-form games. Experiments show that LLMs exhibit consistent strategic reasoning levels across diverse opponents, confirming the framework's robustness and generalization capability. We also analyze the effects of two key mechanisms (Chat Mechanism and Memory Mechanism) on strategic reasoning performance. Results indicate that the Chat Mechanism significantly degrades strategic reasoning, whereas the Memory Mechanism enhances it. These insights position CHBench as a promising tool for evaluating LLM capabilities, with significant potential for future research and practical applications.

cs.AI

Simultaneous All-Pay Auctions with Budget Constraints

The all-pay auction, a classic competitive model, is widely applied in scenarios such as political elections, sports competitions, and research and development, where all participants pay their bids regardless of winning or losing. However, in the traditional all-pay auction, players have no budget constraints, whereas in real-world scenarios, players typically face budget constraints. This paper studies the Nash equilibrium of two players with budget constraints across multiple heterogeneous items in a complete-information framework. The main contributions are as follows: (1) a comprehensive characterization of the Nash equilibrium in single-item auctions with asymmetric budgets and valuations; (2) the construction of a joint distribution Nash equilibrium for the two-item scenario; and (3) the construction of a joint distribution Nash equilibrium for the three-item scenario. Unlike the unconstrained all-pay auction, which always has a Nash equilibrium, a Nash equilibrium may not exist when players have budget constraints. Our findings highlight the intricate effects of budget constraints on bidding strategies, providing new perspectives and methodologies for theoretical analysis and practical applications of all-pay auctions.

econ.TH

Optimal Calibrated Signaling in Digital Auctions

In digital advertising, online platforms allocate ad impressions through real-time auctions, where advertisers typically rely on autobidding agents to optimize bids on their behalf. Unlike traditional auctions for physical goods, the value of an ad impression is uncertain and depends on the unknown click-through rate (CTR). While platforms can estimate CTRs more accurately using proprietary machine learning algorithms, these estimates/algorithms remain opaque to advertisers. This information asymmetry naturally raises the following questions: how can platforms disclose information in a way that is both credible and revenue-optimal? We address these questions through calibrated signaling, where each prior-free bidder receives a private signal that truthfully reflects the conditional expected CTR of the ad impression. Such signals are trustworthy and allow bidders to form unbiased value estimates, even without access to the platform's internal algorithms. We study the design of platform-optimal calibrated signaling in the context of second-price auction. Our first main result fully characterizes the structure of the optimal calibrated signaling, which can also be computed efficiently. We show that this signaling can extract the full surplus -- or even exceed it -- depending on a specific market condition. Our second main result is an FPTAS for computing an approximately optimal calibrated signaling that satisfies an IR condition. Our main technical contributions are: a reformulation of the platform's problem as a two-stage optimization problem that involves optimal transport subject to calibration feasibility constraints on the bidders' marginal bid distributions; and a novel correlation plan that constructs the optimal distribution over second-highest bids.

cs.GT

Ex-Ante Truthful Distribution-Reporting Mechanisms

This paper studies mechanism design for revenue maximization in a distribution-reporting setting, where the auctioneer does not know the buyers' true value distributions. Instead, each buyer reports and commits to a bid distribution in the ex-ante stage, which the auctioneer uses as input to the mechanism. Buyers strategically decide the reported distributions to maximize ex-ante utility, potentially deviating from their value distributions. As shown in previous work, classical prior-dependent mechanisms such as the Myerson auction fail to elicit truthful value distributions at the ex-ante stage, despite satisfying Bayesian incentive compatibility at the interim stage. We study the design of ex-ante incentive compatible mechanisms, and aim to maximize revenue in a prior-independent approximation framework. We introduce a family of threshold-augmented mechanisms, which ensures ex-ante incentive compatibility while boosting revenue through ex-ante thresholds. Based on these mechanisms, we construct the Peer-Max Mechanism, which achieves an either-or approximation guarantee for general non-identical distributions. Specifically, for any value distributions, its expected revenue either achieves a constant fraction of the optimal social welfare, or surpasses the second-price revenue by a constant fraction, where the constants depend on the number of buyers and a tunable parameter. We also provide an upper bound on the revenue achievable by any ex-ante incentive compatible mechanism, matching our lower bound up to a constant factor. Finally, we extend our approach to a setting where multiple units of identical items are sold to buyers with multi-unit demands.

cs.GT

Static Segmentation by Tracking: A Label-Efficient Approach for Fine-Grained Specimen Image Segmentation

We study image segmentation in the biological domain, particularly trait segmentation from specimen images (e.g., butterfly wing stripes, beetle elytra). This fine-grained task is crucial for understanding the biology of organisms, but it traditionally requires manually annotating segmentation masks for hundreds of images per species, making it highly labor-intensive. To address this challenge, we propose a label-efficient approach, Static Segmentation by Tracking (SST), based on a key insight: while specimens of the same species exhibit natural variation, the traits of interest show up consistently. This motivates us to concatenate specimen images into a ``pseudo-video'' and reframe trait segmentation as a tracking problem. Specifically, SST generates masks for unlabeled images by propagating annotated or predicted masks from the ``pseudo-preceding'' images. Built upon recent video segmentation models, such as Segment Anything Model 2, SST achieves high-quality trait segmentation with only one labeled image per species, marking a breakthrough in specimen image analysis. To further enhance segmentation quality, we introduce a cycle-consistent loss for fine-tuning, again requiring only one labeled image. Additionally, we demonstrate the broader potential of SST, including one-shot instance segmentation in natural images and trait-based image retrieval.

cs.CV