Skip to main content
arXiv is now an independent nonprofit! Learn more

Showing 1–50 of 213 results for author: Xin, Z

.
  1. arXiv:2608.14360  [pdf, ps, other

    math.RT

    On Geometric Models of String Algebras: Uniqueness of Surfaces and Existence of Red Punctures

    Authors: Zheng Xin, Lingchun Zhang

    Abstract: A geometric model for string algebras was recently established in \cite{BC24}. Building upon this framework, we characterize the class of string algebras whose geometric models are unique up to equivalence of labelled tiled surfaces. Moreover, we provide a necessary and sufficient condition for all geometric models of a string algebra to be entirely free of red punctures, and further give a comb… ▽ More

    Submitted 14 August, 2026; originally announced August 2026.

    Comments: 37 pages, 24 Figures

    MSC Class: 05E10; 16G20; 05C10

  2. arXiv:2608.12719  [pdf, ps, other

    cs.GT cs.AI

    Error-Aware Reverse Auction Mechanism for Large Language Model Routing

    Authors: Haolong Chen, Zhengyuan Xin, Liang Zhang, Lei Xue, Guangxu Zhu

    Abstract: Routing each query to a cost-effective large language model (LLM) is critical for balancing quality and cost, yet most routers rely on a centralized task center to predict model performance, creating an information-risk mismatch and a scalability bottleneck as the model pool grows. We propose a market-based routing paradigm that shifts ex-ante prediction to LLM providers via a reverse auction, whe… ▽ More

    Submitted 12 August, 2026; originally announced August 2026.

  3. arXiv:2607.27278  [pdf, ps, other

    cs.CV

    OVEarth-Bench: Evaluating Category Breadth and Query Diversity for Open-Vocabulary Earth Observation

    Authors: Kaiyu Li, Zepeng Xin, Zixuan Jiang, Jing Fu, Lanxuan Xue, Lingyu Zhang, Xiangyong Cao

    Abstract: Open-vocabulary Earth observation (EO) aims to localize geospatial concepts specified in natural language rather than a fixed label set. Existing benchmarks, however, usually cover narrow category vocabularies or limited query forms. To fill this gap, we introduce OVEarth-Bench, which extends existing evaluation in two directions: category breadth, through broad hierarchical category coverage with… ▽ More

    Submitted 2 August, 2026; v1 submitted 29 July, 2026; originally announced July 2026.

  4. arXiv:2607.19962  [pdf, ps, other

    cs.AI

    EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization

    Authors: Xinbang Dai, Zheyu Xin, Huikang Hu, Lin Ren, Rihui Jin, Guohui Xiao, Guilin Qi, Kuicai Dong, Zhaocheng Du, Yuyang Zhang

    Abstract: Large Reasoning Models (LRMs) often suffer from overthinking due to redundant verification steps. Existing approaches for mitigating overthinking, such as fast-slow thinking switching and reasoning trajectory compression, fail to make a fine-grained distinction between beneficial and redundant steps within the LRM's reasoning process, and may thus impair reasoning capability in their pursuit of ef… ▽ More

    Submitted 22 July, 2026; originally announced July 2026.

    Comments: 9 pages, 7 figures, accepted by IJCAI 2026

  5. arXiv:2607.18153  [pdf, ps, other

    cs.CV

    Robust Multimodal Dynamic Object Segmentation

    Authors: Zhe Xin, Hanzhi Chang, Penghui Huang, Yinian Mao, Guoquan Huang

    Abstract: Dynamic object segmentation plays a critical role in many visual applications such as static scene reconstruction from dynamic videos. However, existing optical flow-based methods fail to ensure consistent static/dynamic segmentation along object boundaries, while 3D reconstruction-based approaches are highly sensitive to reconstruction errors. To address these limitations, we present a dynamic ob… ▽ More

    Submitted 20 July, 2026; originally announced July 2026.

    Comments: Accepted by IEEE International Conference on Robotics & Automation ICRA 2026

  6. arXiv:2607.17115  [pdf, ps, other

    math.AP

    Enhanced stability and asymptotic limits to the non-isentropic compressible fluid-particle interaction model with thermal effects

    Authors: Fucai Li, Jinkai Ni, Zhouping Xin

    Abstract: In Einstein's seminal work [Ann. Physik, 17 (1905), 549-560], he pointed out that the temperature of a fluid influences the motion of suspended particles dramatically. To describe the effect of the temperature in this physical process more precisely, Boudin et al. [ESAIM Proc., 28 (2009), 195-210] introduced a new fluid-particle interaction model containing of the non-isentropic compressible Euler… ▽ More

    Submitted 19 July, 2026; originally announced July 2026.

    Comments: 76 pages

    MSC Class: 35Q30; 76N10; 35Q83; 35B40

  7. arXiv:2607.15865  [pdf, ps, other

    cs.CL

    An MLIR-Based Compilation Method for Large Language Models

    Authors: Pengchao Hu, Zhibin Xin, Yifan Chen, Yangyang Zhou, Liang Wang, Xin Zhang

    Abstract: Large Language Models (LLMs) have become the dominant workload on modern AI accelerators, yet deploying them on specialized hardware still faces two core challenges: how to import a trained model into a compiler-friendly intermediate representation, and how to efficiently schedule the autoregressive inference loop under limited on-chip memory. This paper presents an MLIR (Multi-Level Intermediate… ▽ More

    Submitted 25 July, 2026; v1 submitted 17 July, 2026; originally announced July 2026.

  8. arXiv:2607.11552  [pdf, ps, other

    math.RT math.RA

    Global dimension of a string algebra

    Authors: Zheng Xin, Lingchun Zhang

    Abstract: In this paper, we characterize the global dimension of a string algebra by using combinatorial methods. Moreover, we establish a necessary and sufficient condition for when the global dimension of a string algebra is infinite.

    Submitted 20 July, 2026; v1 submitted 13 July, 2026; originally announced July 2026.

    Comments: 13 pages, to appear in Arch. Math. (Basel)

    MSC Class: 16G10; 16E10; 18G20

  9. arXiv:2606.29977  [pdf, ps, other

    eess.IV cs.CV cs.LG

    A multi-architecture study of specificity refinement and false-positive mechanism analysis in prostate MRI

    Authors: Yongbo Shu, Kewen Chen, Yifeng Yuan, Zirui Xin, Luo Lei, Yang Yang, Xi Chen, Aijing Luo

    Abstract: Objectives: To characterize residual false positives in prostate MRI detection, and to evaluate a lightweight post-hoc refinement head for case-level specificity. Materials and Methods: This retrospective study used PI-CAI (5-fold cross-validation) and Prostate158 (n=158; external). A context-aware evidence head and an 89,216-parameter refinement head were trained on a frozen detection backbone; t… ▽ More

    Submitted 29 June, 2026; originally announced June 2026.

    Comments: 29 pages, 6 figures, 5 tables

  10. arXiv:2606.25877  [pdf, ps, other

    cs.RO

    TacVerse: A Multi-Sensor Dataset and Benchmark for Cross-Sensor Vision-Based Tactile Perception

    Authors: Lan Wei, Gurmeher Khurana, Sirine Bhouri, Wenhao Hong, Zeyuan Xin, Qingzheng Cong, Wen Fan, Yanzheng Xiang, Dandan Zhang

    Abstract: Vision-based tactile sensors (VBTSs) enable robots to infer contact geometry and force-related cues by imaging deformation through an internal camera, yet generalisation across sensor designs remains poorly understood. We present TacVerse, a multi-sensor dataset and benchmark for cross-sensor vision-based tactile perception. The dataset contains 106,800 tactile images from seven VBTSs and supports… ▽ More

    Submitted 24 June, 2026; originally announced June 2026.

  11. arXiv:2606.07462  [pdf, ps, other

    cs.AI

    Act As a Real Researcher: A Suite of Benchmarks Evaluating Frontier LLMs and Agentic Harnesses in Research Lifecycle

    Authors: Jiayu Wang, Weijiang Lv, Bowen Fu, Jing Fu, Jiayi Song, Lingyu Zhang, Lanxuan Xue, Luodi Chen, Zepeng Xin, Kaiyu Li, Xiangyong Cao

    Abstract: As foundation models advance and agent scaffolding becomes increasingly sophisticated, agents have demonstrated remarkable proficiency in complex, long-horizon coding tasks and even autonomous experiment execution. Despite their evolution from research assistants into autonomous research agents, these systems still exhibit significant limitations in field sensitivity, research ethics, and nuanced… ▽ More

    Submitted 5 June, 2026; originally announced June 2026.

  12. arXiv:2605.23719  [pdf, ps, other

    cs.CV cs.AI

    Weierstrass Positional Encoding for Vision Transformers

    Authors: Zhihang Xin, Rui Wang, Xitong Hu, Xiaojun Wu

    Abstract: Vision Transformers have achieved remarkable success in computer vision, but their common use of learnable one-dimensional positional encodings weakens the inherent two-dimensional spatial structure of images after patch flattening. Existing positional encodings often lack geometric constraints and do not preserve a monotonic relationship between Euclidean spatial distances and sequential index di… ▽ More

    Submitted 20 May, 2026; originally announced May 2026.

  13. arXiv:2605.20035  [pdf, ps, other

    cs.CV

    Stage-adaptive Token Selection for Efficient Omni-modal LLMs

    Authors: Zijie Xin, Jie Yang, Ruixiang Zhao, Tianyi Wang, Fengyun Rao, Jing Lyu, Xirong Li

    Abstract: Omni-modal large language models (om-LLMs) achieve unified audio-visual understanding by encoding video and audio into temporally aligned token sequences interleaved at the window level. However, processing these dense non-textual tokens throughout the LLM incurs substantial computational overhead. Although training-free token selection can reduce this cost, existing methods either focus on visual… ▽ More

    Submitted 19 May, 2026; originally announced May 2026.

    Comments: Code Link: https://github.com/xxayt/SEATS

  14. arXiv:2605.18577  [pdf, ps, other

    cs.CV

    OmniPro: A Comprehensive Benchmark for Omni-Proactive Streaming Video Understanding

    Authors: Ruixiang Zhao, Jie Yang, Zijie Xin, Tianyi Wang, Fengyun Rao, Jing LYU, Xirong Li

    Abstract: Omni-proactive streaming video understanding, i.e., autonomously deciding when to speak and what to say from continuous audio-visual streams, is an emerging capability of omni-modal large language models. Existing benchmarks fall short in three key aspects: they rely primarily on visual signals, adopt polling or fixed-timestamp protocols instead of true proactive evaluation, and cover only a limit… ▽ More

    Submitted 18 May, 2026; originally announced May 2026.

    Comments: Project page: https://ruixiangzhao.github.io/OmniPro

  15. arXiv:2605.18498  [pdf

    cs.LG cs.AI

    DBES: A Systematic Benchmark and Metric Suite for Evaluating Expert Specialization in Large-Scale MoEs

    Authors: Jing Wang, Hongxuan Lu, Jazze Young, Shu Wang, Zhimin Xin

    Abstract: Expert specialization in Mixture-of-Experts (MoE) models remains poorly understood, with traditional evaluations conflating architectural load-balancing with functional specialization. We introduce DBES, a comprehensive diagnostic framework combining a multi-domain benchmark with five theoretically grounded metrics: Routing Specialization, Normalized Effective Rank, Domain Isolation, Routing Stiff… ▽ More

    Submitted 18 May, 2026; originally announced May 2026.

  16. arXiv:2605.00923  [pdf

    eess.IV cs.CV

    A Proof-of-Concept Study of Multitask Learning for Cranial Synthetic CT Generation Across Heterogeneous MRI Field Strengths

    Authors: Zhuoyao Xin, Yiren Zhang, Christopher Wu, Dong Liu, Chunming Gu, Elena Greco, Erik H. Middlebrooks, Jun Hua, Jia Guo

    Abstract: Accurate synthesis of computed tomography (CT) images from magnetic resonance imaging (MRI) is clinically valuable for cranial applications such as attenuation correction, radiotherapy planning, and image-guided interventions. However, heterogeneity across MRI field strengths and acquisition protocols limits the generalizability of existing methods. In this study, we formulate cranial CT synthesis… ▽ More

    Submitted 30 April, 2026; originally announced May 2026.

    Comments: Published in Medical Physics (2026). DOI: 10.1002/mp.70429

    Journal ref: Medical Physics, 53(5): e70429, 2026

  17. arXiv:2604.11998  [pdf, ps, other

    cs.CV cs.AI

    The Second Challenge on Cross-Domain Few-Shot Object Detection at NTIRE 2026: Methods and Results

    Authors: Xingyu Qiu, Yuqian Fu, Jiawei Geng, Bin Ren, Jiancheng Pan, Zongwei Wu, Hao Tang, Yanwei Fu, Radu Timofte, Nicu Sebe, Mohamed Elhoseiny, Lingyi Hong, Mingxi Cheng, Xingqi He, Runze Li, Xingdong Sheng, Wenqiang Zhang, Jiacong Liu, Shu Luo, Yikai Qin, Yaze Zhao, Yongwei Jiang, Yixiong Zou, Zhe Zhang, Yang Yang , et al. (49 additional authors not shown)

    Abstract: Cross-domain few-shot object detection (CD-FSOD) remains a challenging problem for existing object detectors and few-shot learning approaches, particularly when generalizing across distinct domains. As part of NTIRE 2026, we hosted the second CD-FSOD Challenge to systematically evaluate and promote progress in detecting objects in unseen target domains under limited annotation conditions. The chal… ▽ More

    Submitted 13 April, 2026; originally announced April 2026.

    Comments: accepted by CVPRW 26 @ NTIRE

  18. arXiv:2604.10702  [pdf, ps, other

    cs.CV cs.AI

    Backbone-Conditional Behavior of Modality Gating in Multi-Modal Prostate MRI Segmentation: A 5-Fold Cross-Validation and Gate Mechanism Analysis

    Authors: Yongbo Shu, Wenzhao Xie, Shanhu Yao, Zirui Xin, Luo Lei, Kewen Chen, Aijing Luo

    Abstract: Robust segmentation of clinically significant prostate cancer (csPCa) on multi-parametric MRI must tolerate frequent degradation of its most informative diffusion sequences. Multi-modal fusion commonly employs learned modality gating under the assumption that gates implement per-sample modality quality routing -- rarely tested directly. We ask how gating behaves across backbone architectures. We s… ▽ More

    Submitted 24 June, 2026; v1 submitted 12 April, 2026; originally announced April 2026.

    Comments: Major revision. Single-fold analysis replaced by 5-fold cross-validation (180 trained models) plus a direct gate-mechanism analysis; conclusions updated to show that modality gating is backbone-conditional. Supersedes v1

  19. arXiv:2604.08322  [pdf, ps, other

    cs.CV

    Fundus-R1: Training a Fundus-Reading MLLM with Knowledge-Aware Reasoning on Public Data

    Authors: Yuchuan Deng, Qijie Wei, Kaiheng Qian, Jiazhen Liu, Zijie Xin, Bangxiang Lan, Jingyu Liu, Jianfeng Dong, Xirong Li

    Abstract: Fundus imaging such as CFP, OCT and UWF is crucial for the early detection of retinal anomalies and diseases. Fundus image understanding, due to its knowledge-intensive nature, poses a challenging vision-language task. An emerging approach to addressing the task is to post-train a generic multimodal large language model (MLLM), either by supervised finetuning (SFT) or by reinforcement learning wit… ▽ More

    Submitted 9 April, 2026; originally announced April 2026.

  20. arXiv:2603.17670  [pdf, ps, other

    cs.RO

    AgentVLN: Towards Agentic Vision-and-Language Navigation

    Authors: Zihao Xin, Wentong Li, Yixuan Jiang, Ziyuan Huang, Bin Wang, Piji Li, Jianke Zhu, Jie Qin, Shengjun Huang

    Abstract: Vision-and-Language Navigation (VLN) requires an embodied agent to ground complex natural-language instructions into long-horizon navigation in unseen environments. While Vision-Language Models (VLMs) offer strong 2D semantic understanding, current VLN systems remain constrained by limited spatial perception, 2D-3D representation mismatch, and monocular scale ambiguity. In this paper, we propose A… ▽ More

    Submitted 18 March, 2026; originally announced March 2026.

    Comments: 19pages, 4 figures

  21. arXiv:2603.13133  [pdf, ps, other

    cs.RO

    DecoVLN: Decoupling Observation, Reasoning, and Correction for Vision-and-Language Navigation

    Authors: Zihao Xin, Wentong Li, Yixuan Jiang, Bin Wang, Runmin Cong, Jie Qin, Shengjun Huang

    Abstract: Vision-and-Language Navigation (VLN) requires agents to follow long-horizon instructions and navigate complex 3D environments. However, existing approaches face two major challenges: constructing an effective long-term memory bank and overcoming the compounding errors problem. To address these issues, we propose DecoVLN, an effective framework designed for robust streaming perception and closed-lo… ▽ More

    Submitted 26 March, 2026; v1 submitted 13 March, 2026; originally announced March 2026.

    Comments: 16 pages, 8 figures, CVPR2026

  22. arXiv:2603.08957  [pdf, ps, other

    cs.MS cs.AI cs.DB

    Automated Tensor-Relational Decomposition for Large-Scale Sparse Tensor Computation

    Authors: Yuxin Tang, Zhiyuan Xin, Zhimin Ding, Xinyu Yao, Daniel Bourgeois, Tirthak Patel, Chris Jermaine

    Abstract: A \emph{tensor-relational} computation is a relational computation where individual tuples carry vectors, matrices, or higher-dimensional arrays. An advantage of tensor-relational computation is that the overall computation can be executed on top of a relational system, inheriting the system's ability to automatically handle very large inputs with high levels of sparsity while high-performance ker… ▽ More

    Submitted 9 March, 2026; originally announced March 2026.

  23. arXiv:2603.08224  [pdf, ps, other

    cs.CV

    SAVE: Speech-Aware Video Representation Learning for Video-Text Retrieval

    Authors: Ruixiang Zhao, Zhihao Xu, Bangxiang Lan, Zijie Xin, Jingyu Liu, Xirong Li

    Abstract: For video-text retrieval, the use of CLIP has been a de facto choice. Since CLIP provides only image and text encoders, this consensus has led to a biased paradigm that entirely ignores the sound track of videos. While several attempts have been made to reintroduce audio -- typically by incorporating an audio encoder and fusing its output with visual features -- these methods face two challenges:… ▽ More

    Submitted 10 March, 2026; v1 submitted 9 March, 2026; originally announced March 2026.

    Comments: Accepted to CVPR2026

  24. Design and characterization of W-band and D-band calibration sources for the AliCPT-1 experiment

    Authors: Xu-Fang Li, Cong-Zhan Liu, Ai-Mei Zhang, Zheng-Wei Li, Xue-Feng Lu, Zhong-Xue Xin, Guo-Feng Wang, Yong-Ping Li, Yong-Jie Zhang, Shi-Bo Shu, Yi-Fei Zhang, Ya-Qiong Li, Zhi Chang, Dai-Kang Yan

    Abstract: Ali Cosmic Microwave Background Polarization Telescope (AliCPT-1) is the first Chinese cosmic microwave background experiment aiming to make sensitive polarization maps of the potential B-mode signal from inflationary gravitational waves. The telescope was deployed on the Tibet Ali site at 5250 m above sea level in early 2025. Before and after each observation season, the instrument performance mu… ▽ More

    Submitted 12 February, 2026; originally announced February 2026.

    Comments: 13 pages, 10 figures

    Journal ref: Journal of Astronomical Telescopes,Instruments, and Systems, 2026, Vol. 12

  25. arXiv:2602.01876  [pdf, ps, other

    math.NA

    PINN-Based Kolmogorov-Arnold Networks with RAR-D Adaptive Sampling for Solving Elliptic Interface Problems

    Authors: Zijuan Xin, Chenyao Wang, Feng Shi, Yizhong Sun

    Abstract: Physics-Informed Neural Networks (PINNs) have become a popular and powerful framework for solving partial differential equations (PDEs), leveraging neural networks to approximate solutions while embedding PDE constraints, boundary conditions, and interface jump conditions directly into the loss function. However, most existing PINN approaches are based on multilayer perceptrons (MLPs), which may r… ▽ More

    Submitted 2 February, 2026; originally announced February 2026.

  26. arXiv:2512.20013  [pdf, ps, other

    cs.CV

    SegEarth-R2: Towards Comprehensive Language-guided Segmentation for Remote Sensing Images

    Authors: Zepeng Xin, Kaiyu Li, Luodi Chen, Wanchen Li, Yuchen Xiao, Hui Qiao, Weizhan Zhang, Deyu Meng, Xiangyong Cao

    Abstract: Effectively grounding complex language to pixels in remote sensing (RS) images is a critical challenge for applications like disaster response and environmental monitoring. Current models can parse simple, single-target commands but fail when presented with complex geospatial scenarios, e.g., segmenting objects at various granularities, executing multi-target instructions, and interpreting implici… ▽ More

    Submitted 22 December, 2025; originally announced December 2025.

  27. arXiv:2512.06433  [pdf, ps, other

    physics.flu-dyn

    Model of incompressible turbulent flows via a kinetic theory

    Authors: Ziyang Xin, Zhaoli Guo, Hudong Chen

    Abstract: Kinetic theory offers a promising alternative to conventional turbulence modelling by providing a mesoscopic perspective that naturally captures non-equilibrium physics such as non-Newtonian effects. In this work, we present an extension and theoretical analysis of the recent kinetic model for incompressible turbulent flows developed by Chen et al. (Atmos. 14(7), 1109, 2023), constructed for unbou… ▽ More

    Submitted 12 March, 2026; v1 submitted 6 December, 2025; originally announced December 2025.

    Comments: 36 page;12 figures

  28. arXiv:2511.16147  [pdf, ps, other

    cs.CL cs.AI

    TS-PEFT: Unveiling Token-Level Redundancy in Parameter-Efficient Fine-Tuning

    Authors: Dabiao Ma, Ziming Dai, Zhimin Xin, Shu Wang, Jian Yang, Haojun Fei

    Abstract: Current Parameter-Efficient Fine-Tuning (PEFT) methods typically operate under an implicit assumption: Once a target module is selected, every token passing through it contributes equally to the downstream task and requires a parameter update. In this paper, we challenge this convention by revealing a pervasive token-level redundancy in the fine-tuning of large models (LMs). We propose TS-PEFT, a… ▽ More

    Submitted 29 January, 2026; v1 submitted 20 November, 2025; originally announced November 2025.

    Comments: 11 pages, 3 figures

  29. arXiv:2511.14593  [pdf, ps, other

    hep-ex

    First measurement of reactor neutrino oscillations at JUNO

    Authors: Angel Abusleme, Thomas Adam, Kai Adamowicz, David Adey, Shakeel Ahmad, Rizwan Ahmed, Timo Ahola, Sebastiano Aiello, Fengpeng An, Guangpeng An, Costas Andreopoulos, Giuseppe Andronico, João Pedro Athayde Marcondes de André, Nikolay Anfimov, Vito Antonelli, Tatiana Antoshkina, Burin Asavapibhop, Didier Auguste, Margherita Buizza Avanzini, Andrej Babic, Jingzhi Bai, Weidong Bai, Nikita Balashov, Roberto Barbera, Andrea Barresi , et al. (1114 additional authors not shown)

    Abstract: Neutrino oscillations, a quantum effect manifesting at macroscopic scales, are governed by lepton flavor mixing angles and neutrino mass-squared differences that are fundamental parameters of particle physics, representing phenomena beyond the Standard Model. Precision measurements of these parameters are essential for testing the completeness of the three-flavor framework, determining the mass or… ▽ More

    Submitted 18 November, 2025; originally announced November 2025.

    Comments: 30 pages, 11 figures

  30. arXiv:2511.14590  [pdf, ps, other

    hep-ex physics.ins-det

    Initial performance results of the JUNO detector

    Authors: Angel Abusleme, Thomas Adam, Kai Adamowicz, David Adey, Shakeel Ahmad, Rizwan Ahmed, Timo Ahola, Sebastiano Aiello, Fengpeng An, Guangpeng An, Costas Andreopoulos, Giuseppe Andronico, João Pedro Athayde Marcondes de André, Nikolay Anfimov, Vito Antonelli, Tatiana Antoshkina, Burin Asavapibhop, Didier Auguste, Margherita Buizza Avanzini, Andrej Babic, Jingzhi Bai, Weidong Bai, Nikita Balashov, Roberto Barbera, Andrea Barresi , et al. (1114 additional authors not shown)

    Abstract: The Jiangmen Underground Neutrino Observatory (JUNO) started physics data taking on 26 August 2025. JUNO consists of a 20-kton liquid scintillator central detector, surrounded by a 35 kton water pool serving as a Cherenkov veto, and almost 1000 m$^2$ of plastic scintillator veto on top. The detector is located in a shallow underground laboratory with an overburden of 1800 m.w.e. This paper present… ▽ More

    Submitted 18 November, 2025; originally announced November 2025.

    Comments: 38 pages, 23 figures

  31. arXiv:2511.07227  [pdf, ps, other

    hep-ex physics.geo-ph

    Prospects for geoneutrino detection with JUNO

    Authors: Thomas Adam, Shakeel Ahmad, Rizwan Ahmed, Fengpeng An, João Pedro Athayde Marcondes de André, Costas Andreopoulos, Giuseppe Andronico, Nikolay Anfimov, Vito Antonelli, Tatiana Antoshkina, Didier Auguste, Marcel Büchner, Weidong Bai, Nikita Balashov, Andrea Barresi, Davide Basilico, Eric Baussan, Marco Beretta, Antonio Bergnoli, Nikita Bessonov, Daniel Bick, Lukas Bieger, Svetlana Biktemerova, Thilo Birkenfeld, Simon Blyth , et al. (605 additional authors not shown)

    Abstract: Geoneutrinos, which are antineutrinos emitted during the decay of long-lived radioactive elements inside Earth, serve as a unique tool for studying the composition and heat budget of our planet. The Jiangmen Underground Neutrino Observatory (JUNO) experiment in China, which has recently completed construction, is expected to collect a sample comparable in size to the entire existing world geoneutr… ▽ More

    Submitted 10 November, 2025; originally announced November 2025.

    Comments: 32 pages, with 13 figures and 5 tables

  32. arXiv:2511.06956  [pdf, ps, other

    astro-ph.IM

    Mock Observations for the CSST Mission: Main Surveys--the Stray Light

    Authors: Xian Jing-Tian, Lin Lin, Fang Yue-Dong, Zhang Xin, Xu You-Hua, Meng Xian-Min, Tian Hao, Zhang Tian-Yi, Ban Zhang, Li Guo-Liang, Xu Shu-Yan, Wang Wei

    Abstract: Stray light significantly influences the detection capabilities of astronomical telescopes. The actual stray-light level during observations depends not only on the telescope's inherent stray-light suppression capability but also on its operational orbit conditions. Accurate estimation of stray-light levels is crucial for assessing image quality and performing realistic scientific simulations. To… ▽ More

    Submitted 10 November, 2025; originally announced November 2025.

  33. arXiv:2510.06616  [pdf, ps, other

    physics.ins-det hep-ex

    Design, waterproofing, and mass production of the 3-inch PMT frontend system of JUNO

    Authors: Jilei Xu, Miao He, Cédric Cerna, Yongbo Huang, Thomas Adam, Shakeel Ahmad, Rizwan Ahmed, Fengpeng An, Costas Andreopoulos, Giuseppe Andronico, João Pedro Athayde Marcondes de André, Nikolay Anfimov, Vito Antonelli, Tatiana Antoshkina, Didier Auguste, Weidong Bai, Nikita Balashov, Andrea Barresi, Davide Basilico, Eric Baussan, Marco Beretta, Antonio Bergnoli, Nikita Bessonov, Daniel Bick, Lukas Bieger , et al. (609 additional authors not shown)

    Abstract: Over 25,600 3-inch photomultiplier tubes (PMTs) have been instrumented for the central detector of the Jiangmen Underground Neutrino Observatory. Each PMT is equipped with a high-voltage divider and a frontend cable with waterproof sealing. Groups of sixteen PMTs are connected to the underwater frontend readout electronics via specialized multi-channel waterproof connectors. This paper outlines th… ▽ More

    Submitted 22 January, 2026; v1 submitted 7 October, 2025; originally announced October 2025.

  34. arXiv:2510.00788  [pdf, ps, other

    cond-mat.mtrl-sci

    Efficient E(3)-equivariant framework for universal charge density prediction

    Authors: Xiwen Li, Zaizhou Xin, Hongyu Yu, Yang Zhong, Xingao Gong, Hongjun Xiang

    Abstract: Electronic structure is ubiquitously obtained via density functional theory (DFT), where the charge density plays a central role. This work presents EdenGNN (Equivariant Density Graph Neural Network), a machine learning (ML) charge density model for electronic structure. Current universal ML charge density models are hampered by prohibitive computational costs. Furthermore, despite being trained o… ▽ More

    Submitted 13 March, 2026; v1 submitted 1 October, 2025; originally announced October 2025.

  35. arXiv:2509.25830  [pdf

    physics.optics

    Integrated Silicon Photonic Multichannel Optical Hybrid for Broadband Parallel Coherent Reception

    Authors: Tong Lin, Yan Fan, Jiao Zhang, Wenqi Yu, Zhengyu Guo, Liu Li, Zhigang Xin, Mingzheng Lei, Ziyang Xiong, Haoran Wang, Hao Deng, Min Zhu, Shihua Chen, Junpeng Lu, Zhenhua Ni

    Abstract: We design and demonstrate a monolithically integrated silicon photonic multichannel optical hybrid for versatile broadband coherent reception, addressing the critical limitations of current wavelength multiplexed systems in scalability and power efficiency. The device combines a phase-compensated 90-degree optical hybrid with four robust three-stage Mach-Zehnder interferometer lattice filters, ena… ▽ More

    Submitted 30 September, 2025; originally announced September 2025.

    Comments: 18 pages, 4 figures

  36. arXiv:2509.10957  [pdf, ps, other

    cs.HC

    The Digital Landscape of God: Narrative, Visuals and Viewer Engagement of Religious Videos on YouTube

    Authors: Rongyi Chen, Ziyan Xin, Qing Xiao, Ruiwei Xiao, Jingjia Xiao, Bingbing Zhang, Hong Shen, Zhicong Lu

    Abstract: The digital transformation of religious practice has reshaped how billions of people engage with spiritual content, with video-sharing platforms becoming central to contemporary religious communication. Yet HCI research lacks systematic understanding of how narrative and visual elements create meaningful spiritual experiences and foster viewer engagement. We present a mixed-methods study of religi… ▽ More

    Submitted 10 April, 2026; v1 submitted 13 September, 2025; originally announced September 2025.

    Comments: This work has been accepted by ICWSM 2026

  37. arXiv:2508.19167  [pdf, ps, other

    cs.CV

    Beyond flattening: a geometrically principled positional encoding for vision transformers with Weierstrass elliptic functions

    Authors: Zhihang Xin, Xitong Hu, Rui Wang

    Abstract: Vision Transformers have demonstrated remarkable success in computer vision tasks, yet their reliance on learnable one-dimensional positional embeddings fundamentally disrupts the inherent two-dimensional spatial structure of images through patch flattening procedures. Traditional positional encoding approaches lack geometric constraints and fail to establish monotonic correspondence between Eucli… ▽ More

    Submitted 26 August, 2025; originally announced August 2025.

  38. arXiv:2508.12247  [pdf, ps, other

    cs.LG cs.AI

    STM3: Mixture of Multiscale Mamba for Long-Term Spatio-Temporal Time-Series Prediction

    Authors: Haolong Chen, Liang Zhang, Zhengyuan Xin, Guangxu Zhu

    Abstract: Recently, spatio-temporal time-series prediction has developed rapidly, yet existing deep learning methods struggle with learning complex long-term spatio-temporal dependencies efficiently. The long-term spatio-temporal dependency learning brings two new challenges: 1) The long-term temporal sequence naturally includes multiscale information, which is hard to extract efficiently; 2) The multiscale… ▽ More

    Submitted 22 May, 2026; v1 submitted 17 August, 2025; originally announced August 2025.

    Comments: Accepted by KDD 2026

  39. arXiv:2508.02340  [pdf, ps, other

    cs.CV cs.IR cs.MM

    Learning Partially-Decorrelated Common Spaces for Ad-hoc Video Search

    Authors: Fan Hu, Zijie Xin, Xirong Li

    Abstract: Ad-hoc Video Search (AVS) involves using a textual query to search for multiple relevant videos in a large collection of unlabeled short videos. The main challenge of AVS is the visual diversity of relevant videos. A simple query such as "Find shots of a man and a woman dancing together indoors" can span a multitude of environments, from brightly lit halls and shadowy bars to dance scenes in black… ▽ More

    Submitted 4 August, 2025; originally announced August 2025.

    Comments: Accepted by ACMMM2025

  40. arXiv:2507.15308  [pdf, ps, other

    cs.CV

    Few-Shot Object Detection via Spatial-Channel State Space Model

    Authors: Zhimeng Xin, Tianxu Wu, Yixiong Zou, Shiming Chen, Dingjie Fu, Xinge You

    Abstract: Due to the limited training samples in few-shot object detection (FSOD), we observe that current methods may struggle to accurately extract effective features from each channel. Specifically, this issue manifests in two aspects: i) channels with high weights may not necessarily be effective, and ii) channels with low weights may still hold significant value. To handle this problem, we consider uti… ▽ More

    Submitted 21 July, 2025; originally announced July 2025.

  41. arXiv:2507.03009  [pdf, ps, other

    cs.CL cs.IR cs.LG

    PDFMathTranslate: Scientific Document Translation Preserving Layouts

    Authors: Rongxin Ouyang, Chang Chu, Zhikuang Xin, Xiangyao Ma

    Abstract: Language barriers in scientific documents hinder the diffusion and development of science and technologies. However, prior efforts in translating such documents largely overlooked the information in layouts. To bridge the gap, we introduce PDFMathTranslate, the world's first open-source software for translating scientific documents while preserving layouts. Leveraging the most recent advances in l… ▽ More

    Submitted 22 September, 2025; v1 submitted 2 July, 2025; originally announced July 2025.

    Comments: 7 pages, 4 figures, EMNLP 2025 System Demonstration

    MSC Class: 68T50; 68T45; 68U10; 68U15 ACM Class: D.2.2; I.2.10; I.2.7; J.0

  42. arXiv:2506.19884  [pdf, ps, other

    cs.OS cs.AI cs.PF cs.SE

    MNN-AECS: Energy Optimization for LLM Decoding on Mobile Devices via Adaptive Core Selection

    Authors: Zhengxiang Huang, Chaoyue Niu, Zhaode Wang, Jiarui Xue, Hanming Zhang, Yugang Wang, Zewei Xin, Xiaotang Jiang, Chengfei Lv, Fan Wu, Guihai Chen

    Abstract: As the demand for on-device Large Language Model (LLM) inference grows, energy efficiency has become a major concern, especially for battery-limited mobile devices. Our analysis shows that the memory-bound LLM decode phase dominates energy use, and yet most existing works focus on accelerating the prefill phase, neglecting energy concerns. We introduce Adaptive Energy-Centric Core Selection (AECS)… ▽ More

    Submitted 24 June, 2025; originally announced June 2025.

  43. A Novel ViDAR Device With Visual Inertial Encoder Odometry and Reinforcement Learning-Based Active SLAM Method

    Authors: Zhanhua Xin, Zhihao Wang, Shenghao Zhang, Wanchao Chi, Yan Meng, Shihan Kong, Yan Xiong, Chong Zhang, Yuzhen Liu, Junzhi Yu

    Abstract: In the field of multi-sensor fusion for simultaneous localization and mapping (SLAM), monocular cameras and IMUs are widely used to build simple and effective visual-inertial systems. However, limited research has explored the integration of motor-encoder devices to enhance SLAM performance. By incorporating such devices, it is possible to significantly improve active capability and field of view… ▽ More

    Submitted 16 June, 2025; originally announced June 2025.

    Comments: 12 pages, 13 figures

    MSC Class: 93C85 ACM Class: I.4

    Journal ref: IEEE Transactions on Industrial Informatics, pp. 1-12, 2025

  44. arXiv:2506.06137  [pdf, ps, other

    cs.LG cs.CL

    Table-r1: Self-supervised and Reinforcement Learning for Program-based Table Reasoning in Small Language Models

    Authors: Rihui Jin, Zheyu Xin, Xing Xie, Zuoyi Li, Guilin Qi, Yongrui Chen, Xinbang Dai, Tongtong Wu, Gholamreza Haffari

    Abstract: Table reasoning (TR) requires structured reasoning over semi-structured tabular data and remains challenging, particularly for small language models (SLMs, e.g., LLaMA-8B) due to their limited capacity compared to large LMs (LLMs, e.g., GPT-4o). To narrow this gap, we explore program-based TR (P-TR), which circumvents key limitations of text-based TR (T-TR), notably in numerical reasoning, by gene… ▽ More

    Submitted 6 June, 2025; originally announced June 2025.

  45. arXiv:2505.09915  [pdf, ps, other

    cs.CV cs.RO

    Large-Scale Gaussian Splatting SLAM

    Authors: Zhe Xin, Chenyang Wu, Penghui Huang, Yanyong Zhang, Yinian Mao, Guoquan Huang

    Abstract: The recently developed Neural Radiance Fields (NeRF) and 3D Gaussian Splatting (3DGS) have shown encouraging and impressive results for visual SLAM. However, most representative methods require RGBD sensors and are only available for indoor environments. The robustness of reconstruction in large-scale outdoor scenarios remains unexplored. This paper introduces a large-scale 3DGS-based visual SLAM… ▽ More

    Submitted 14 May, 2025; originally announced May 2025.

  46. arXiv:2505.09296  [pdf, ps, other

    math.AP

    Global well-posedness of the Cauchy problem for the modified Whitham equations

    Authors: Han Cui, Yuexun Wang, Zhouping Xin

    Abstract: This paper aims to show global existence and modified scattering for the solutions of the Cauchy problem to the modified Whitham equations for small, smooth and localized initial data. The main difficulties come from slow decay and non-homogeneity of the Fourier multiplier $(\sqrt{\tanh ξ/ξ})ξ$, which will be overcome by introducing an interaction multiplier theorem and estimating the weighted nor… ▽ More

    Submitted 14 May, 2025; originally announced May 2025.

  47. arXiv:2504.14585  [pdf, ps, other

    astro-ph.IM physics.ins-det

    Development of 6-inch 80-170 GHz broadband silicon plated horn antenna arrays for primordial gravitational wave search

    Authors: Yuanhang He, Shibo Shu, Yaqiong Li, Xuefeng Lu, Ye Chai, Xiang Li, Zhi Chang, He Gao, Yudong Gu, Xufang Li, Zhengwei Li, Zhouhui Liu, Guofeng Wang, Zhongxue Xin, Daikang Yan, Aimei Zhang, Yifei Zhang, Yongjie Zhang, Wenhua Shi, Juexian Cao, Congzhan Liu

    Abstract: Searching for primordial gravitational wave in cosmic microwave background (CMB) polarization signal is one of the key topics in modern cosmology. Cutting-edge CMB telescopes requires thousands of pixels to maximize mapping speed. Using modular design, the telescope focal plane is simplified as several detector modules. Each module has hundreds of pixels including antenna arrays, detector arrays,… ▽ More

    Submitted 20 April, 2025; originally announced April 2025.

    Comments: 13 pages, 14 figures, to be published in Research in Astronomy and Astrophysics

  48. arXiv:2504.09644  [pdf, other

    cs.CV

    SegEarth-R1: Geospatial Pixel Reasoning via Large Language Model

    Authors: Kaiyu Li, Zepeng Xin, Li Pang, Chao Pang, Yupeng Deng, Jing Yao, Guisong Xia, Deyu Meng, Zhi Wang, Xiangyong Cao

    Abstract: Remote sensing has become critical for understanding environmental dynamics, urban planning, and disaster management. However, traditional remote sensing workflows often rely on explicit segmentation or detection methods, which struggle to handle complex, implicit queries that require reasoning over spatial context, domain knowledge, and implicit user intent. Motivated by this, we introduce a new… ▽ More

    Submitted 13 April, 2025; originally announced April 2025.

  49. arXiv:2504.07435  [pdf, other

    cs.GT

    Opportunity-Cost-Driven Reward Mechanisms for Crowd-Sourced Computing Platforms

    Authors: Shuhao Zheng, Ziyue Xin, Zonglun Li, Xue Liu

    Abstract: This paper introduces a game-theoretic model tailored for reward distribution on crowd-sourced computing platforms. It explores a repeated game framework where miners, as computation providers, decide their computation power contribution in each round, guided by the platform's designed reward distribution mechanism. The reward for each miner in every round is based on the platform's randomized tas… ▽ More

    Submitted 10 April, 2025; originally announced April 2025.

    Comments: 10 pages, 1 figure, accepted as FULL paper in IEEE International Conference on Blockchain and Cryptocurrency 2025 (ICBC'2025)

  50. arXiv:2503.19351  [pdf, ps, other

    cs.CV

    Multi-Object Sketch Animation by Scene Decomposition and Motion Planning

    Authors: Jingyu Liu, Zijie Xin, Yuhan Fu, Ruixiang Zhao, Bangxiang Lan, Xirong Li

    Abstract: Sketch animation, which brings static sketches to life by generating dynamic video sequences, has found widespread applications in GIF design, cartoon production, and daily entertainment. While current methods for sketch animation perform well in single-object sketch animation, they struggle in multi-object scenarios. By analyzing their failures, we identify two major challenges of transitioning f… ▽ More

    Submitted 2 August, 2025; v1 submitted 25 March, 2025; originally announced March 2025.

    Comments: Accepted by ICCV 2025