arXiv ScienceSearch

arXiv subjects

Qixiang Wang

Publications and source records attributed to Qixiang Wang.

4 recordsLinked to original sources

Evaluating the Performance of Large Language Models on GAOKAO Benchmark

Large Language Models(LLMs) have demonstrated remarkable performance across various natural language processing tasks; however, how to comprehensively and accurately assess their performance becomes an urgent issue to be addressed. This paper introduces GAOKAO-Bench, an intuitive benchmark that employs questions from the Chinese GAOKAO examination as test samples, including both subjective and objective questions. To align with human examination methods, we design a method based on zero-shot settings to evaluate the performance of LLMs. With human evaluation, we obtain the converted total score of LLMs, including GPT-4, ChatGPT and ERNIE-Bot.Our findings reveal that LLMs have achieved competitive scores in Chinese GAOKAO examination, while they exhibit significant performance disparities across various subjects. We also use LLMs to grade the subjective questions, and find that model scores achieve a moderate level of consistency with human scores. In conclusion, this research contributes a robust evaluation benchmark for future large language models and offers valuable insights into the advantages and limitations of such models.

cs.CL

JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency

We introduce JoyAI-LLM Flash, an efficient Mixture-of-Experts (MoE) language model designed to redefine the trade-off between strong performance and token efficiency in the sub-50B parameter regime. JoyAI-LLM Flash is pretrained on a massive corpus of 20 trillion tokens and further optimized through a rigorous post-training pipeline, including supervised fine-tuning (SFT), Direct Preference Optimization (DPO), and large-scale reinforcement learning (RL) across diverse environments. To improve token efficiency, JoyAI-LLM Flash strategically balances \emph{thinking} and \emph{non-thinking} cognitive modes and introduces FiberPO, a novel RL algorithm inspired by fibration theory that decomposes trust-region maintenance into global and local components, providing unified multi-scale stability control for LLM policy optimization. To enhance architectural sparsity, the model comprises 48B total parameters while activating only 2.7B parameters per forward pass, achieving a substantially higher sparsity ratio than contemporary industry leading models of comparable scale. To further improve inference throughput, we adopt a joint training-inference co-design that incorporates dense Multi-Token Prediction (MTP) and Quantization-Aware Training (QAT). We release the checkpoints for both JoyAI-LLM-48B-A3B Base and its post-trained variants on Hugging Face to support the open-source community.

cs.CL

A Note on Hodge theoretic anabelian geometry

Grothendieck's anabelian conjectures predict that certain classes of varieties over number fields are largely determined by their {é}tale fundamental groups. A theorem of Mochizuki shows that for hyperbolic curves over number fields or $p$-adic fields, dominant morphisms bijectively correspond to open homomorphisms between their {é}tale fundamental groups. Motivated by non-abelian Hodge theory, we formulate a Hodge-theoretic version of the anabelian conjecture in which the Galois action is replaced by the natural $\mathbb{C}^\times$-action on the pro-algebraic completion of the fundamental group arising from non-abelian Hodge theory. In particular, we prove a Hodge-theoretic analog of Mochizuki's theorem for smooth projective hyperbolic curves over $\mathbb{C}$. We also obtain a higher-dimensional analogue for complex hyperbolic manifolds of ball quotient type and discuss possible extensions to non-$K(π,1)$ spaces replacing fundamental groups by homotopy types.

math.AG

The relative GAGA Theorem and an application to the analytic mapping stacks

We prove a relative GAGA theorem for perfect and pseudo-coherent complexes in non-archimedean analytic geometry, allowing bases given by Fredholm analytic rings, including those associated from affinoid perfectoid spaces. This answers a question raised in \cite{heuer2024padicnonabelianhodgetheory}. As an application, we show that for a proper scheme \(X\) and an Artin stack \(Y\) with suitable conditions, the analytification of the algebraic mapping stack \(\mathrm{Map}(X,Y)\) agrees with the intrinsic analytic mapping stack \(\mathrm{Map}(X^{\mathrm{an}},Y^{\mathrm{an}})\).

math.AG