arXiv ScienceSearch

EXPLORE CONNECTIONS

LBPE: Long-token-first Tokenization to Improve Large Language Models

Follow the relationships that help you find your next source.

Based on 525 indexed works selected for this snapshot; the candidate window is limited. Counts describe this index, not the complete source archives. Prepared from the PostgreSQL corpus; source versions are checked before display. Snapshot 2026-09-23. Up to 32 works or names per graph.

Subjects & research connections

Connections use shared source subjects, names, places, and explicitly mentioned entities. A shared label is not evidence of a citation, experimental result, or verified species identification.

cs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLcs.CLLBPE: Long-token-first Tokenization to Improve Large Language ModelsLBPE: Long-token-first …RELIC: Retrieving Evidence for Literary Claims2ELITR-Bench: A Meeting Assistant Benchmark for Long-Context Language Models3Diversity-grounded Channel Prototypical Learning for Out-of-Distribution Intent Detection4EndoCogniAgent: Closed-Loop Agentic Reasoning with Self-Consistency Validation for Endoscopic Diagnosis5SafetyFlow: An Agent-Flow System for Automated LLM Safety Benchmarking6Geometric Uncertainty for Detecting and Correcting Hallucinations in LLMs7DartQuant: Efficient Rotational Distribution Calibration for LLM Quantization8Text-only adaptation in LLM-based ASR through text denoising9Calibrated Confidence Expression for Radiology Report Generation10Learning Diagnostic Reasoning for Decision Support in Toxicology11A Survey on Long-Term Memory Security in LLM Agents: Attacks, Defenses, and Governance Across the Memory Lifecycle12GroupTravelBench: Benchmarking LLM Agents on Multi-Person Travel Planning13MONA: Muon Optimizer with Nesterov Acceleration for Scalable Language Model Training14CONCAT: Consensus- and Confidence-Driven Ad Hoc Teaming for Efficient LLM-Based Multi-Agent Systems15Recovering the Zipfian Distribution in Unsupervised Term Discovery16KaLM-Reranker-V1: Fast but Not Late Interaction for Compressed Document Reranking17Explanation-Guided Medical Named Entity Recognition with Stability and Boundary Awareness for Atopic Dermatitis18From Plausible to Actionable: A Position on LLM Self-Explanations19DFAH-Bench: Benchmarking Observable Agent Instability in Financial Decision-Making20Hy-MultiTurn: A Six-Dimensional Benchmark for Deep Multi-Turn Dialogue Understanding21Query-Side Attacks on GNN-Based KGQA: Tracing Failures from Entity Linking to Answer Generation22Compositional Failure in Audio-Visual LLMs: Late-Layer Prior Dominance Under Cross-modal Conflict23AI Writers Have a Consistent Stylometric Footprint, but AI Editors Do Not24Quantitative Evidence Mining for Plausibility-Aware Biomedical AI: A Narrative Review and Conceptual Framework25Lngram v2: Latent N-Gram Memory with Interpretable Discrete Representations26LLM-Anchored Paralinguistic Enrichment for Alzheimer's Disease Detection27Disentangling Topology and Diversity in Multi-Agent LLMs for Multilingual Low-Resource Emotion Detection28Rollback the World, Keep the Reflection: Rollback-Induced Reflection for Long-Horizon LLM Agents29Playing log(N)-Questions over Wikipedia Abstracts: How Per-Round Errors Compound Under Information Asymmetry30PAGE: Partition-Aware Gated KV-Cache Eviction31Beyond Task Completion: Training Capable and Safe Computer-Use Agents32
Relationships as a list (31)

Citation graph

Arrows run from the citing work to its reference. Only explicit source/provider reference lists are used. Incoming links cover this candidate window; this is not a global citation count.

No supported relationships are available in this snapshot. This does not mean that no relationships exist.

Collaborator network

Shared authorship within up to 200 candidate works (1 examined); up to 32 names shown. Names are matched as supplied, without verified person disambiguation. Shared credit does not necessarily establish personal collaboration.

1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared work1 shared workGuiguang DingGuiguang DingHaoran LianHaoran LianHui ChenHui ChenJianwei NiuJianwei NiuPeng LiuPeng LiuShasha MoShasha MoYizhe XiongYizhe XiongZijia LinZijia Lin
Relationships as a list (28)

Publication timeline

Publication years for these 32 related works; 0 have no source publication date. This is a discovery sample, not a measure of research output or growth.

2022 · 1 work
2024 · 2 works
2025 · 2 works
2026 · 27 works

Semantic map

Model: all-minilm. Positions are a two-dimensional approximation of embedding similarity; proximity is not a citation or proof of agreement. Results come from the snapshot’s candidate window.

No current, compatible embeddings are available for related works in this snapshot. A semantic map appears after background embedding and snapshot generation.