Skip to main content
arXiv is now an independent nonprofit! Learn more

Showing 1–50 of 570 results for author: Zou, Z

.
  1. arXiv:2608.21184  [pdf, ps, other

    gr-qc

    Scalar dark matter in space-based gravitational-wave detectors: center-of-mass motion, size breathing, and TDI projection

    Authors: Rui-Yang Hu, Yuan-Zhi Li, An-Qi Wang, Zong-Ru Zou, Fa-Peng Huang, Cheng-Gang Qin

    Abstract: Ultralight scalar dark matter can make space-based gravitational-wave detectors respond through both the scalar charge of freely falling test masses and scalar-induced changes of local solid length scales. Existing space-detector forecasts usually model the former as a center-of-mass force, while ground-based interferometer studies show that scalar fields can also act through material and optical-… ▽ More

    Submitted 21 August, 2026; originally announced August 2026.

    Comments: 16 pages, 5 figures

  2. arXiv:2608.17427  [pdf, ps, other

    cs.CV

    Counterfactual Anatomy-guided Spatial-Temporal Decoding for Annotation-Free Hallucination Mitigation in Medical VLMs

    Authors: Yifan Lu, Adinath Dukre, Abhijit Das, Ziyun Zou, Haolin Yang, Yutong Xie, Imran Razzak

    Abstract: Medical vision-language models (Med-VLMs) have demonstrated strong performance on medical visual question answering, yet they remain prone to hallucination, generating clinically unsupported statements that are insufficiently grounded in image evidence. Mitigation methods applied during decoding offer a practical solution, but they typically lack anatomical awareness or rely heavily on ground trut… ▽ More

    Submitted 18 August, 2026; originally announced August 2026.

    Comments: Accepted by MICCAI 2026

  3. arXiv:2608.14269  [pdf, ps, other

    hep-ph

    Nonstandard Solution for Anomaly Cancellation as Seesaw Neutrino Origin in the SM

    Authors: Zi-Yue Zou, Chia-Wei Liu, Zhong-Lv Huang, Xiao-Gang He

    Abstract: For fixed Standard Model (SM) non-Abelian representations of 15 chiral fermions with arbitrary hypercharges, anomaly cancellation admits the usual assignment and a distinct nonstandard solution. In the latter, the colored exotic quark and exotic lepton weak doublets and exotic lepton singlet have zero hypercharge, whereas the two colored exotic quark singlets carry opposite hypercharges $-q$ and… ▽ More

    Submitted 14 August, 2026; originally announced August 2026.

    Comments: 8 pages, 1 figures

  4. arXiv:2608.13136  [pdf, ps, other

    cs.CL cs.AI cs.DB cs.MA

    LigBench: A Unified and Human-Aligned Benchmark for LLM-based Research Idea Generation

    Authors: Chenrun Wang, Mingxuan Zhu, Tiancheng Huang, Wenjie Li, Yujie Zhang, Zichen Zhu, Zhiying Zou, Kai Yu, Lu Chen

    Abstract: With the rapid advancement of large language models (LLMs), research idea generation has attracted increasing attention. Existing approaches enable LLMs to retrieve relevant literature and propose novel ideas for research areas. However, current evaluation practices for idea generation remain fragmented and lack objective standards, often relying on direct LLM scoring, which limits their ability t… ▽ More

    Submitted 13 August, 2026; originally announced August 2026.

    Comments: 17 pages

  5. arXiv:2608.13030  [pdf, ps, other

    cs.CR cs.MA cs.NI

    InterSAGE: The Secure and Verifiable Interoperability Protocol for An Internet of Agents

    Authors: Zhenhua Zou, Sheng Guo, Qiuyang Zhan, Lepeng Zhao, Shuo Li, Zhuotao Liu

    Abstract: The emerging Internet of Agents enables LLM-powered agents to discover peers, invoke tools, and delegate tasks across organizational boundaries. Existing protocols increasingly define how agents exchange messages, but not how an agent proves its identity, authorization, advertised capabilities, or accountability after delegation. We present InterSAGE, a trust-native protocol suite that supplies th… ▽ More

    Submitted 13 August, 2026; v1 submitted 13 August, 2026; originally announced August 2026.

    Comments: 35 pages, 4 figures, 7 tables. Positioning paper

  6. arXiv:2608.12898  [pdf, ps, other

    cs.CV cs.AI

    NaviDC-OCR: Navigating Document Parsing Across Digital and Camera-Captured Documents

    Authors: Peng Cai, Zhaofan Zou, Shifa Liu, Yikun Wang, Jiawei Tang, Kaicheng Yang, Meng Tong, MingKun Jiang, Zhongjiang He, Hao Sun

    Abstract: Document parsing aims to transform unstructured documents into structured and machine-readable representations. Recent advances in Vision-Language Models (VLMs) have significantly advanced document parsing. However, existing approaches still face two major challenges. First, decoupled VLM-based methods heavily rely on accurate layout analysis, where geometric distortions in camera-captured documen… ▽ More

    Submitted 18 August, 2026; v1 submitted 13 August, 2026; originally announced August 2026.

  7. arXiv:2608.10831  [pdf, ps, other

    quant-ph

    Projection measurement of the comb basis through free-electron-photon interactions

    Authors: Zihang Zou, Feng-Xiao Sun, Yunquan Liu, Qiongyi He

    Abstract: Free electrons, driven by rapid advances in photon-induced near-field electron microscopy, have emerged as a promising platform for quantum information processing, including quantum computing and quantum sensing. However, conventional measurements that rely on the electron energy loss spectrum (EELS) are inherently destructive to electron qubits, thereby constraining their applicability. In this L… ▽ More

    Submitted 11 August, 2026; originally announced August 2026.

  8. arXiv:2608.08025  [pdf, ps, other

    cs.RO cs.AI

    DA-NBV: A Direction-Aware Next-Best-View Planner for Efficient 3D Reconstruction of Ships at Sea

    Authors: Jiaming Chen, Juntao Yang, Zhentao Zou, Qi Ming, Yi Yu, Zhihang Zhong, Xue Yang, Xue Jiang, Yue Zhou

    Abstract: Accurate 3D reconstruction of ships at sea is important for maritime supervision, damage assessment, and autonomous maritime operations. Although 3D reconstruction has advanced considerably, high-quality data acquisition still largely relies on manually designed trajectories or skilled operators, resulting in high costs and limited scalability. Next-best-view (NBV) planning automates this process… ▽ More

    Submitted 8 August, 2026; originally announced August 2026.

    Comments: 16pages, 11 figures

  9. Depth-Guided Video Object Counting in Crowded Scenes

    Authors: Yuanjing Xu, Xinyan Liu, Weidong Chen, Zixuan Zou, Linhao Zhang, Zhuangzhe Meng, Antoni B. Chan, Weigang Zhang

    Abstract: Our primary objective is to advance video object counting in crowded scenes, aiming to robustly count all instances of a target category based on given text or visual prompts. Existing methods rely on RGB information, limiting their discriminative ability in crowded and occluded conditions. To address this, we propose a Depth-Guided Detector (DG-Det) along with a general post-processing pipeline.… ▽ More

    Submitted 6 August, 2026; originally announced August 2026.

    Comments: Accepted at ACM Multimedia 2026

  10. arXiv:2608.05523  [pdf, ps, other

    cs.CV

    HERA: Historical Evidence Routing Adapter for Physical Prediction in Latent World Models

    Authors: Yuanruyi, Yue Cao, Haojia Gao, Guanqiu Guo, Ziyuezhang, Shangqin, Junbo Tan, Bokui Chen, Zhuo Zou, Xueqian Wang

    Abstract: Predictive video models have emerged as promising world models by learning latent visual dynamics from large-scale video. Yet these models remain challenged by physical events under occlusion, where later predictions may depend on object evidence that is no longer available in the current view. Addressing this challenge requires historical evidence not only to be preserved but also to remain acces… ▽ More

    Submitted 5 August, 2026; originally announced August 2026.

  11. arXiv:2608.04527  [pdf, ps, other

    cs.RO

    Retrieve in Time, Correct in Frequency

    Authors: Yuze Fan, Yue Cao, Pengjie Gao, Haojia Gao, Guangqiu Guo, Ziyue Zhang, Junbo Tan, Bokui Chen, Zhuo Zou, Xueqian Wang

    Abstract: Frozen vision-language-action (VLA) policies generate temporally extended action chunks, but long-horizon manipulation remains vulnerable to accumulated execution error and visual aliasing across task stages. Successful rollouts provide useful corrective evidence, yet current frame retrieval can return progress-misaligned actions,while direct replay or time-domain fusion can overwrite the reactive… ▽ More

    Submitted 6 August, 2026; v1 submitted 5 August, 2026; originally announced August 2026.

  12. arXiv:2608.04106  [pdf, ps, other

    cs.CV eess.IV

    LoRetta: A Foundation Model and Extensive Dataset for Global-Scale Remote Sensing Dense Image Matching

    Authors: Siwei Yu, Han Guo, Zhenwei Shi, Zhengxia Zou

    Abstract: Dense image matching establishes pixel-wise correspondences and underpins broad applications in computer vision and photogrammetry. However, extending dense matching to global-scale remote sensing remains challenging because image pairs may differ in acquisition time, season, viewpoint, spatial resolution, and land-cover state. The resulting large geometric offsets, partial overlap, and intrinsica… ▽ More

    Submitted 4 August, 2026; originally announced August 2026.

    Comments: 17 pages, 12 figures, 6 tables. Submitted to IEEE Transactions on Pattern Analysis and Machine Intelligence. Project page: https://siweiyu.com/work/loretta/

  13. arXiv:2608.00440  [pdf, ps, other

    cs.CV

    Poplar: A Scalable Pipeline for Human-Centric Image Dataset Synthesis

    Authors: Zhishan Zou

    Abstract: Recent image generators can synthesize convincing human-centric images, yet producing a useful collection remains different from producing a single successful image. A human-centric dataset must cover varied people and contexts, avoid implausible attribute combinations, preserve an everyday photographic character, and expose quality-control decisions at scale. We present Poplar, a reproducible Spe… ▽ More

    Submitted 1 August, 2026; originally announced August 2026.

    Comments: code: https://github.com/choucisan/poplar website: https://choucisan.github.io/publications/poplar

  14. arXiv:2607.27969  [pdf, ps, other

    cs.CV

    FootprintNet: State-Transition-Guided Dynamic Footprint Learning for Multi-temporal Remote Sensing Change Detection

    Authors: Haotian Zhang, Hao Chen, Han Guo, Zhengxia Zou, Zhenwei Shi

    Abstract: Despite substantial progress in remote sensing multi-temporal change detection (MTCD), most existing MTCD methods still represent the dynamic process at each spatial location over the entire observation period using a single change category associated with the final observation. This implicit single-change assumption limits their ability to characterize regions of recurrent change closely related… ▽ More

    Submitted 30 July, 2026; originally announced July 2026.

  15. arXiv:2607.23114  [pdf, ps, other

    astro-ph.HE

    Diverse Morphologies of GRB X-Ray Plateaus within a Common Magnetar Framework

    Authors: Xiao-Fei Dong, Yong-Feng Huang, Nurimangul Nurmamat, Chen Deng, Ze-Cheng Zou, Fan Xu, Abdusattar Kurban, Chen Du, Chen-Ran Hu, Jin-Jun Geng

    Abstract: The origin of the X-ray plateau phase in gamma-ray bursts (GRBs) remains an open problem. In particular, it is unclear whether GRBs with different temporal morphologies (i.e., with a rising, flat, or decaying plateau) arise from a common underlying mechanism. Although magnetar energy injection is a leading explanation, previous studies have primarily inferred magnetar properties on a burst-by-burs… ▽ More

    Submitted 25 July, 2026; originally announced July 2026.

    Comments: 14 pages, 7 figures, 1 table. Submitted. Comments are welcome

  16. arXiv:2607.22662  [pdf, ps, other

    cs.AI

    CuraWeb: Joint Optimization of Quality, Redundancy, and Diversity for Web-Scale Pretraining Data

    Authors: Peiguang Li, Yongwei Zhou, Juncheng Diao, Yuchun Fan, Jian Yang, Jianxiao Yang, Zhongda Su, Shuguang Jiao, Xiao Wei, Zhiye Zou, Gan Dong, Zhizhao Zeng, Rongxiang Weng, Jingang Wang, Xunliang Cai

    Abstract: Open-web corpora curated via highly selective filters, such as FineWeb-Edu and DCLM, constitute the core of LLM pretraining data and have significantly advanced LLM performance. However, these pipelines typically rely on singular optimization objectives, which inevitably narrows distributional diversity and marginalizes long-tail knowledge, thereby restricting data coverage and underutilizing the… ▽ More

    Submitted 28 June, 2026; originally announced July 2026.

  17. arXiv:2607.22166  [pdf, ps, other

    cs.RO cs.AI

    Learning Spatiotemporal Decision Priors for Efficient Path Planning under Partial Observability

    Authors: Yi Liu, Hongda Zhang, Leyao Zou, Chunlei Meng, Ziqing Zhou, Yuning Chen, Zhuo Zou, Lida Xu, Zhongxue Gan, Chun Ouyang

    Abstract: Path planning under partial observability remains challenging because an agent must make long-horizon navigation decisions from only locally bounded observations. Nevertheless, historical trajectories contain reusable experience-guided directional preferences. Classical planners, however, typically solve each instance from scratch and lack an explicit mechanism to exploit such transferable decisio… ▽ More

    Submitted 24 July, 2026; originally announced July 2026.

  18. arXiv:2607.19228  [pdf, ps, other

    cs.CV

    IGGT4D: Streaming 4D Instance-Grounded Geometry Transformer

    Authors: Zhengyu Zou, Hao Li, Kuixuan Jiao, Liu Liu, Tingyang Xiao, Xiaolin Zhou, Fangzhou Hong, Zhizhong Su, Dingwen Zhang, Ziwei Liu

    Abstract: Real-world spatial intelligence requires agents to understand scenes from continuous video streams, where objects move, persist, disappear, and reappear over time. While recent spatial foundation models have enabled generalizable feed-forward 3D reconstruction, most streaming methods remain geometry-centric and lack temporally consistent object-level understanding. Meanwhile, existing semantic rec… ▽ More

    Submitted 21 July, 2026; originally announced July 2026.

    Comments: Project Page: https://iggt4d.github.io

  19. arXiv:2607.18848  [pdf, ps, other

    quant-ph

    Quantum-Enhanced Multi-Objective Optimization

    Authors: Maolin Luo, Jiapei Zhuang, Zuoheng Zou, Man-Hong Yung

    Abstract: Multi-objective combinatorial optimization requires identifying Pareto-optimal trade-off solutions among conflicting objectives, often making it more demanding than its single-objective counterpart. Although quantum multi-objective optimization methods have begun to emerge, most existing quantum optimization workflows are still built around single-objective or fixed-scalarization settings. Buildin… ▽ More

    Submitted 21 July, 2026; originally announced July 2026.

    Comments: 20 pages, 13 figures

  20. arXiv:2607.18846  [pdf, ps, other

    cs.DS cs.CR

    Private Approximation of Graph Spectra and Cuts via Spectral Amplifiers

    Authors: Chenglin Fan, Jingcheng Liu, Pan Peng, Hangyu Xu, Zongrui Zou

    Abstract: We study the problem of releasing a synthetic graph that approximates the sizes of all cuts of an input graph under edge-level differential privacy. If one insists on purely additive error, the optimal worst-case error is $\widetildeΘ(n^{3/2})$. If one allows a small multiplicative slack, an information-theoretic exponential-time mechanism achieves nearly linear additive error, but the best known… ▽ More

    Submitted 21 July, 2026; originally announced July 2026.

    Comments: 77 pages

  21. arXiv:2607.15561  [pdf, ps, other

    math.NA

    RCLUPPr: a new randomized CholeskyQR with LU preconditioning

    Authors: Haoran Guan, Zhenyu Zou, Yufeng Wei, Yipei Chen, Peiting You, Yuwei Fan

    Abstract: In this work, we present the comprehensive rounding error analysis of RCLUPPr proposed in \cite{RCLUPP}, which is a novel randomized CholeskyQR-type algorithm performing LU decomposition with partial pivoting (LUPP decomposition) directly on the tall-skinny $X\in\mathbb{R}^{m\times n}$ with $m \ge n$ and $\mbox{rank}(X)=n$. In contrast to the existing RCLUPP in \cite{RCLUPP}, which applies matrix… ▽ More

    Submitted 16 July, 2026; originally announced July 2026.

    MSC Class: 15A23; 65F25; 65F30; 65G50

  22. Nexus: Native Mesh Generation with Diffusion

    Authors: Hanxiao Wang, Ying-Tian Liu, Yuan-Chen Guo, Qi-Yuan Feng, Zi-Xin Zou, Ding Liang, Biao Zhang, Yan-Pei Cao

    Abstract: Generating high-quality triangle meshes is essential for film, gaming, and interactive 3D applications. Mainstream methods rely on mesh serialization and autoregressive processes, which stuggles in effective inference and is sensitive to error accumulation. In this paper, we present Nexus, a diffusion method that achieves holistic mesh generation via decoupled vertex and topology generation. First… ▽ More

    Submitted 15 July, 2026; originally announced July 2026.

  23. arXiv:2607.13301  [pdf, ps, other

    quant-ph cond-mat.mes-hall cond-mat.mtrl-sci

    Precision quantum simulation of magnon spectra and interactions

    Authors: Trond I. Andersen, Nikita Astrakhantsev, Jeronimo Martinez, Will Morong, Johannes Motruk, Dario Rossi, Brayden Ware, Bryce Kobrin, Weijie Wu, Elizabeth Bennewitz, Manuel Rudolph, Tom Westerhout, Amira Abbas, Rajeev Acharya, Laleh Aghababaie Beni, Ross Alcaraz, Sayra Alcaraz, Markus Ansmann, Frank Arute, Kunal Arya, Walt Askew, Juan Atalaya, Christopher Ayala, Ryan Babbush, Brian Ballard , et al. (307 additional authors not shown)

    Abstract: Quantum simulation promises to advance materials discovery by accurately simulating complex states of matter, their microscopic excitations, and macroscopic response functions. The central challenge in resolving the underlying interacting dynamics is to combine high-fidelity evolution with the sophisticated control necessary to manipulate individual quasi-particles in quantum many-body states. Her… ▽ More

    Submitted 14 July, 2026; originally announced July 2026.

  24. arXiv:2607.13101  [pdf, ps, other

    cs.LG cs.AI

    TSSM: Triaxial State Space Model for Global Station Weather Forecasting with Temporal-Variable-Historical Modeling

    Authors: Songru Yang, Zili Liu, Tao Han, Ben Fei, Fenghua Ling, Lei Bai, Chang Liu, Xiangyang Ji, Zhenwei Shi, Zhengxia Zou

    Abstract: Global Station Weather Forecasting (GSWF) is pivotal for localized and extreme weather prediction over key regions. Despite efforts to exploit look-back windows, existing methods show limited accuracy gains and struggle with extreme events and error accumulation. These limitations stem from overreliance on short-term patterns, which are insufficient to capture chaotic weather dynamics, especially… ▽ More

    Submitted 14 July, 2026; originally announced July 2026.

  25. arXiv:2607.12477  [pdf, ps, other

    cs.CV

    Self in Space: Benchmarking Self-Awareness and Spatial Cognition in UAV Embodied Intelligence

    Authors: Zhishan Zou, Guoyan Sun, Zhiwei Wei, Jiancheng Pan, Yujie Li, Mugen Peng, Wenjia Xu

    Abstract: Autonomous UAV systems increasingly rely on multimodal large language models (MLLMs) to operate in complex real-world environments. Such embodied scenarios require not only understanding the surrounding space but also maintaining a coherent representation of the agent itself. However, existing UAV-oriented approaches and benchmarks remain largely environment-centric, primarily focusing on spatial… ▽ More

    Submitted 14 July, 2026; v1 submitted 14 July, 2026; originally announced July 2026.

    Comments: Website:https://choucisan.github.io/publications/self-in-space ; Code:https://github.com/IntelliSensing/Self-in-Space

  26. arXiv:2607.10553  [pdf, ps, other

    cs.RO

    SLIDER: Sparse History-Guided Aerial Robot Target Search using Sliding Local Maps

    Authors: Xiaolei Hou, Zheng Pan, Hua Lan, Zhenghao Zou, Yinhong Chen, Chenxi Zhu, Yang Lyu, Jinwen Hu, Chunhui Zhao

    Abstract: Efficient exploration and target search in large-scale unknown environments remain challenging for aerial robots due to the demands of broad spatial coverage, fine-grained perception, and real-time decision-making. This paper presents SLIDER, a lightweight and memory-efficient framework that avoids reliance on globally dense maps by combining a local sliding map with sparse global history informat… ▽ More

    Submitted 11 July, 2026; originally announced July 2026.

    Comments: Accepted by IEEE Robotics and Automation Letters (RA-L), 2026. https://github.com/Poaos/SLIDER

  27. arXiv:2607.10165  [pdf, ps, other

    cs.CV cs.AI

    EmoStyle: Affective Conditioning of Style-Specialist Experts for Emotional Image Generation

    Authors: Dexiang Hong, Yijie Guo, Weidong Chen, Xinyan Liu, Zixuan Zou, Zhendong Mao, Yongdong Zhang

    Abstract: Emotion-aware artistic image generation requires an image to match the input prompt, follow the specified artistic style, and convey the target emotion. In this challenge, the main difficulty is that the visual and affective attributes available in the training data are not explicitly provided at test time. Without these attributes, the generator has to decide not only what to depict, but also how… ▽ More

    Submitted 11 July, 2026; originally announced July 2026.

  28. arXiv:2607.04302  [pdf, ps, other

    cs.LG cs.AI cs.AR cs.CL cs.PF

    HiFA4: Training-Free 4-bit FlashAttention on Ascend HIF4 NPUs for LLM Inference

    Authors: Hui Dong, Yanzhao Li, Jie Gao, Chunlu Li, Zhiyuan Zhang, Yupeng Sun, Zhenyuan Chen, Zhiqiang Zou

    Abstract: We present HiFA4, a post-training operator-level design that executes both QK^T and PV in FlashAttention as 4-bit HIF4 Cube GEMMs for LLM inference on Ascend NPUs, while maintaining the online softmax state in FP16. To our knowledge, HiFA4 is the first Ascend-HIF4-targeted design of this kind evaluated on standard NLP benchmarks. HiFA4 combines two mechanisms. Smooth-QK applies a calibration-sta… ▽ More

    Submitted 5 July, 2026; originally announced July 2026.

    Comments: 22 pages

  29. arXiv:2607.02876  [pdf, ps, other

    hep-ph hep-ex

    Revisiting $\bar B^0 \rightarrow Λ_c^+ \bar p$ decay with higher twist corrections

    Authors: Zhou Rui, Zhi-Tian Zou, Ying Li

    Abstract: We investigate the single-charmed baryonic decays $\bar B^0 \to Λ_c^+ \bar p$ and $\bar B^0 \to \barΛ_c^- p$, which receive contributions from both $W$-emission and $W$-exchange topologies, within the framework of perturbative QCD (PQCD). Higher-power corrections associated with the hadronic light-cone distribution amplitudes (LCDAs) of both the initial- and final-state hadrons are systematically… ▽ More

    Submitted 2 July, 2026; originally announced July 2026.

    Comments: 21 pages,2 figures

  30. arXiv:2606.29879  [pdf, ps, other

    cs.CV cs.AI

    LWDrive: Layer-Wise World-Model-Guided Vision-Language Model Planning for Autonomous Driving

    Authors: Chen Yang, Yuhao Wei, Ze Xu, Ziheng Zou, Shuang Liang, Delin Ouyang, Lingfeng Qi, Jie Li, Guofa Li

    Abstract: Vision-Language Models (VLMs) provide powerful semantic understanding and commonsense reasoning for End-to-End Autonomous Driving (E2E-AD) planning. However, trajectories directly generated by VLMs often encode only coarse driving intentions and remain insufficient for geometrically accurate, future-aware, and multi-view-grounded planning. To address these limitations, we develop the Layer-Wise Wo… ▽ More

    Submitted 30 June, 2026; v1 submitted 29 June, 2026; originally announced June 2026.

  31. State Space Models Meet Remote Sensing: A Survey

    Authors: Qinzhe Yang, Chenyang Liu, Jia Xu, Zhenwei Shi, Zhengxia Zou

    Abstract: State Space Models (SSMs), designed for long-range modeling, offer linear computational complexity and strong capabilities in capturing long-range dependencies. In the field of remote sensing, SSMs have gained popularity due to their effectiveness in addressing unique challenges such as dense visual predictions, multi-modal remote sensing data, and temporal remote sensing data, which have also yie… ▽ More

    Submitted 23 June, 2026; originally announced June 2026.

    Comments: 25 pages, 5 figures, has been published in SCIS SCIQ1 IF=8.1 https://doi.org/10.1007/s11432-025-4780-1

  32. Efficient Remote Sensing Instance Segmentation with Linear-Time State Space Distilled Visual Foundation Models

    Authors: Qinzhe Yang, Keyan Chen, Jia Xu, Zhenwei Shi, Zhengxia Zou

    Abstract: The computational complexity of Transformers scales quadratically with the number of tokens, which significantly constrains the efficiency of vision models, particularly recent ViT-based foundation models in dense prediction tasks. Instance segmentation, a typical dense visual prediction task in the remote sensing field, faces similar challenges. In this paper, inspired by the recent advances of k… ▽ More

    Submitted 23 June, 2026; originally announced June 2026.

    Comments: 17 pages, 11 figures, has been published in IEEE TGRS vol. 64, pp. 5625417-5625417, 2026, Art no. 5625417, doi: 10.1109/TGRS.2026.3696104

    Journal ref: IEEE Transactions on Geoscience and Remote Sensing, vol. 64, pp. 5625417-5625417, 2026, Art no. 5625417

  33. arXiv:2606.25312  [pdf, ps, other

    cs.CV

    LEVIRDet: A Million-Scale 159-Category Dataset and Foundation Model for Universal Remote Sensing Object Detection

    Authors: Qinzhe Yang, Dongyu Wang, Haohan Niu, Jia Xu, Zhenwei Shi, Zhengxia Zou

    Abstract: Remote sensing object detection has advanced rapidly with the development of large-scale benchmarks and modern detection architectures. However, existing datasets and detectors remain fragmented. Most benchmarks focus on limited categories, fixed spatial resolutions, or a single sensor, while detectors still struggle to work across different sensors and categorical systems. In this paper, we intro… ▽ More

    Submitted 4 July, 2026; v1 submitted 23 June, 2026; originally announced June 2026.

    Comments: 18 pages, 9 figures

  34. arXiv:2606.22088  [pdf, ps, other

    math.GT math.DG

    Isometric free finite group actions on non-positively curved 3-manifolds

    Authors: Zhengyu Zou

    Abstract: Let $M$ be a closed orientable $3$-manifold admitting a metric of nonpositive sectional curvature (an NPC metric), and let $G$ be a finite group acting freely on $M$ by orientation-preserving diffeomorphisms. Previous results showed that $M$ admits a $G$-invariant NPC metric except possibly when $M$ is a graph manifold. In this paper, we resolve the remaining case by proving that $M$ also admits a… ▽ More

    Submitted 14 July, 2026; v1 submitted 20 June, 2026; originally announced June 2026.

    Comments: 29 pages, 3 figures. This version includes a new statement of the metric extension criterion using the language of charges of framed Seifert fibered spaces

    MSC Class: 57M60 (Primary) 53C21; 57K35; 57M50 (Secondary)

  35. arXiv:2606.16317  [pdf, ps, other

    cs.CV

    Training-free sparse attention based on cumulative energy filtering

    Authors: Chunlu Li, Yixuan Pan, Bai Du, Zhenyuan Chen, Yanzhao Li, Hui Dong, Hui Wang, Zhiqiang Zou

    Abstract: Sparse attention accelerates Diffusion Transformers (DiTs) for video generation by computing only the important tokens while skipping the rest. The token selection strategy is key to balancing sparsity and accuracy. We formulate the token filtering process as a dual-goal optimization problem: maximizing sparsity and minimizing accuracy degradation. Existing algorithms cannot fulfill both objective… ▽ More

    Submitted 15 June, 2026; originally announced June 2026.

  36. arXiv:2606.16219  [pdf, ps, other

    cs.CE cs.LG physics.comp-ph

    Graphical conditional generative modeling for digital twin modeling

    Authors: Zongren Zou, Théo Bourdais, Ricardo Baptista, Houman Owhadi

    Abstract: Digital twin modeling, including control and data assimilation under model uncertainty, often faces an open-ended fidelity problem: adding variables, data streams, and time scales can indefinitely increase model complexity, ultimately producing systems that are difficult to maintain, validate, interpret, and use for stress or safety testing. As an alternative, one can seek parsimonious stochastic… ▽ More

    Submitted 15 June, 2026; originally announced June 2026.

  37. arXiv:2606.15822  [pdf, ps, other

    cs.AI cs.CR

    TrustedARI: Towards Trust-Native Agentic Routing Infrastructure for Agentic AI

    Authors: Qi Li, Zhenhua Zou, Shuo Li, Mingwei Xu, Zhuotao Liu

    Abstract: AI agents increasingly access external models, tools, and services through Agentic Routing Infrastructure (ARI) to manage the overhead of heterogeneous interfaces and fragmented subscriptions. Yet, the architecture of ARI introduces fundamental trust risks: it obtains plaintext access to agent queries and service responses, while leaving agents unable to verify that their queries are routed to int… ▽ More

    Submitted 14 June, 2026; originally announced June 2026.

    ACM Class: I.2.0

  38. arXiv:2606.11532  [pdf, ps, other

    cs.CR

    Hiding the Trees in the Forest: Building Network Covert Channels with Hash-Based Covert Carrier Filtering

    Authors: Zexiao Zou, Zhiqiang Wang, Baoxu Liu, Yuyang Han, Yan Zhang

    Abstract: As an effective anti-censorship mechanism, network covert channels can provide data privacy protection and ensure communication security. However, the covertness of existing network covert channels primarily depends on the secrecy of their covert algorithms. With the increasing depth of research in this field, the difficulty of breaking such algorithms has gradually decreased. Once the algorithm i… ▽ More

    Submitted 9 June, 2026; originally announced June 2026.

  39. arXiv:2606.10014  [pdf, ps, other

    astro-ph.HE astro-ph.SR

    X-rays breaking out of pre-explosion ejecta mark a supernova's first light

    Authors: Weimin Yuan, Qiu-Ju Huang, Jin-Ping Zhu, Yun-Wei Yu, Dong Xu, Chen Zhang, Zhuo Li, Yuan Liu, Tao An, Giulia Gianfagna, Weikang Zheng, Guowang Du, Xing Liu, Ji-An Jiang, Johan P. U. Fynbo, Alexei S. Pozanenko, Junjie Jin, Yi Yang, Jinsong Deng, Hui Sun, Guang-Lei Wu, Yu-Hao Zhang, Bao Wang, Yu Wang, Xiangyu Wang , et al. (108 additional authors not shown)

    Abstract: Massive stars die as core-collapse supernovae, whose optical light emerges days after the implosion. Theory predicts that the initial collapse-driven shock, upon breaking through the star and dense circumstellar medium, emits a brief thermal flash of soft X-rays and ultraviolet. Yet these elusive first signals have remained largely undetected, owing to limited wide-field soft X-ray monitoring. Her… ▽ More

    Submitted 9 July, 2026; v1 submitted 8 June, 2026; originally announced June 2026.

    Comments: 8 figures, 5 tables

  40. arXiv:2606.05671  [pdf, ps, other

    cs.CL

    QueryAgent-R1: Bridging Query Generation and Product Retrieval for E-Commerce Query Recommendation

    Authors: Dike Sun, Zheng Zou, Jingtong Zang, Qi Sun, Huaipeng Zhaoand Tao Luo, Xiaoyi Zeng

    Abstract: Query recommendation in e-commerce search aims to proactively suggest queries that match users' potential interests. However, existing methods mainly optimize query-level relevance, while neglecting whether the retrieved products align with users' downstream preferences. This mismatch often leads to high query click through rates (CTR) but low product conversion rates (CVR). To bridge this gap, we… ▽ More

    Submitted 3 June, 2026; originally announced June 2026.

  41. arXiv:2606.03676  [pdf, ps, other

    quant-ph

    Macroscopic Spin GHZ States with a Levitated Ferromagnet

    Authors: Xueqi Ni, Zhixing Zou, Ping Koy Lam, Tao Wang, Jiangbin Gong

    Abstract: The generation of macroscopic quantum states can drive both fundamental physics and quantum technologies. This work proposes a top-down approach to the generation of macroscopic spin GHZ states using a levitated ferromagnet, where a strong locking between the collective spin and the lattice rotation enables mechanical control of the collective spin. We quantify the metrological advantage of the re… ▽ More

    Submitted 2 June, 2026; originally announced June 2026.

  42. arXiv:2606.01178  [pdf, ps, other

    hep-ph

    Probing the imaginary parts and their $q^2$ dependences for the tau $g-2$ and EDM

    Authors: Xin-Yu Du, Xiao-Gang He, Zhong-Lv Huang, Chia-Wei Liu, Zi-Yue Zou

    Abstract: The $τ$ anomalous magnetic dipole moment (MDM) $a_τ= (g-2)_τ/2$ and electric dipole moment (EDM) $d_τ$, are precision probes of electroweak dynamics and possible new physics sources, yet both remain weakly constrained experimentally. Treated as generalized form factors, these quantities exhibit a generic $q^2$ dependence for an off-shell interacting photon. For timelike momentum transfer above the… ▽ More

    Submitted 10 June, 2026; v1 submitted 31 May, 2026; originally announced June 2026.

    Comments: 18 pages, 6 figures, 2 tables

  43. arXiv:2605.29002  [pdf, ps, other

    cs.LG cs.DC

    FedQHD: Closed-Form Function-Space Federated Reinforcement Learning

    Authors: Yuchen Hou, Yongshan Chen, Zhuowen Zou, Calvin Yeung, Mohsen Imani, Tian Lan, Mahdi Imani

    Abstract: Federated reinforcement learning enables decentralized agents to collaboratively improve policies or value estimates without exchanging raw trajectories. However, FedAvg-style parameter averaging is not function-space consistent: when clients use heterogeneous encoders or even identical nonlinear networks, averaged parameters need not correspond to the weighted average of client value functions in… ▽ More

    Submitted 27 May, 2026; originally announced May 2026.

  44. arXiv:2605.28048  [pdf, ps, other

    cs.RO

    SAFEVPR: Patch-Based Conformal Verification for Safe Cross-Condition Sequence Visual Place Recognition

    Authors: Ha Sier, Jiaqiang Zhang, Zhuo Zou, Xianjia Yu, Tomi Westerlund

    Abstract: Sequence-based visual place recognition (VPR) for SLAM and robot relocalization must decide whether the retrieved top-1 candidate is safe to accept. Conformal prediction is a natural framework for this accept/reject decision, but its finite-sample guarantees rely on exchangeability between calibration and deployment (test) data, which is violated under cross-condition deployment. We introduce SAFE… ▽ More

    Submitted 27 May, 2026; originally announced May 2026.

  45. arXiv:2605.27813  [pdf, ps, other

    cs.CV cs.AI cs.LG

    Residualized Temporal Sparse Autoencoders for Interpreting Diffusion Models

    Authors: Calvin Yeung, Prathyush Poduval, Ali Zakeri, Zhuowen Zou, Mohsen Imani

    Abstract: Text-to-image diffusion models generate images through an iterative denoising process, so internal neural layers produce trajectories of activations rather than single static representations. Sparse autoencoders (SAEs) have recently been used to decompose diffusion activations into interpretable feature directions, but most approaches analyze activations at individual timesteps or condition on tim… ▽ More

    Submitted 26 May, 2026; originally announced May 2026.

  46. arXiv:2605.26553  [pdf, ps, other

    astro-ph.HE

    Constraining the Supernova Remnant Environment of FRB 190520B with Dispersion Measure and Scattering Timescale

    Authors: Jia-Peng Wei, Chen Deng, Gwenael Giacinti, Ze-Cheng Zou, Chen-Ran Hu, Yong-Feng Huang, Jin-Jun Geng

    Abstract: FRB 190520B is a repeating fast radio burst source whose large dispersion measure (DM) and temporal broadening suggest a dense and evolving local environment. In this work, we test the possibility that FRB 190520B originates from the core-collapse of a massive star so that its central engine is embedded in a supernova remnant (SNR) expanding into a wind environment, whose evolution is described by… ▽ More

    Submitted 26 May, 2026; originally announced May 2026.

  47. arXiv:2605.26532  [pdf, ps, other

    stat.ME

    Global Average Treatment Effects for Individualized Randomization Experiments with Aggregate Data

    Authors: Shuguang Yu, Ting Li, Yuchen Lu, Chengchun Shi, Fan Zhou, Zhichao Zou, Peng Zhen, Hongtu Zhu

    Abstract: Individualized randomized experiments are central to online platforms for optimizing personalized decisions in complex environments. In two-sided markets, however, standard treatment effect estimation is often invalid due to strong temporal and cross-unit interference, a challenge compounded when only aggregated data are available because of privacy or system constraints. To address these issues,… ▽ More

    Submitted 26 May, 2026; originally announced May 2026.

  48. arXiv:2605.25737  [pdf, ps, other

    cs.CV

    SFR-Net: Learning Scale-Frustum Representations for Ultra-Wide Area Remote Sensing Image Segmentation

    Authors: Chuyu Zhong, Keyan Chen, Qinzhe Yang, Bowen Chen, Zhengxia Zou, Zhenwei Shi

    Abstract: Pixel count and geographical coverage are two key characteristics of remote sensing images. Existing remote sensing image segmentation methods typically focus on images with either a small pixel count or a large pixel count but limited geographical coverage. In this paper, we introduce a novel segmentation task targeting ultra-wide area (UWA) remote sensing images, characterized by both a large pi… ▽ More

    Submitted 25 May, 2026; originally announced May 2026.

  49. arXiv:2605.25457  [pdf

    cond-mat.soft physics.atom-ph

    Non-equilibrium pathway to mesoscale ordering in ethanol-water binary liquid

    Authors: Xinyue Jiang, Yating Shang, Jianhui Li, Zhaoyong Zou, Yanxia Zuo, Yuqun Xie

    Abstract: Ethanol-water mixtures are a classic example of thermodynamic non-ideality, yet the structural origin of their pronounced anomalies, such as volume contraction and a large negative excess entropy, has remained a long-standing puzzle. Here, we demonstrate these anomalies are not equilibrium properties but calorimetric fingerprint of an arrested phase transition. By imposing periodic thermal oscilla… ▽ More

    Submitted 25 May, 2026; originally announced May 2026.

  50. arXiv:2605.21440  [pdf, ps, other

    cs.CV

    ReMATF: Recurrent Motion-Adaptive Multi-scale Turbulence Mitigation for Dynamic Scenes

    Authors: Zhiming Liu, Zhicheng Zou, Nantheera Anantrasirichai

    Abstract: Atmospheric turbulence severely degrades video quality by introducing distortions such as geometric warping, blur, and temporal flickering, posing significant challenges to both visual clarity and temporal consistency. Current state-of-the-art methods are based on transformer, 3D architectures and require multi-frame input, but their large computational cost and memory usage limit real-time deploy… ▽ More

    Submitted 20 May, 2026; originally announced May 2026.