Skip to main content
arXiv is now an independent nonprofit! Learn more

Showing 1–50 of 290 results for author: Cui, Q

.
  1. arXiv:2608.18764  [pdf, ps, other

    cs.IR

    GateDiffInt: Gate-Mediated Controllable Diffusion and Multi-Intent LLM Distillation for User Behavior Modeling

    Authors: Jialong Duan, Zichen Zhang, Zirui Tu, Zheng Zhang, Zepeng Li, Qingyao Cui, Qinwen Wang, Yudan Liu, Luo Yang, Yao Hu

    Abstract: Existing ranking models encode intent only implicitly, making it hard to disentangle structured intents of varying strength and temporal scale. Noise and intent in behavior sequences are mutually reinforcing---we call this Noise--Intent Coupling (NIC). Noise dilutes true intents, while the lack of structured intent priors leaves denoising without a clear target. To address NIC, we propose GateDiff… ▽ More

    Submitted 19 August, 2026; v1 submitted 19 August, 2026; originally announced August 2026.

  2. arXiv:2608.18613  [pdf, ps, other

    cs.AI cs.CR

    CTIFoundry: An Agent-Native Corpus Scaffold for Cyber Threat Intelligence

    Authors: Yutong Cheng, Changze Li, Qian Cui, Wei Ding, Lingzhi Wang, Yan Chen, Peng Gao

    Abstract: Cyber threat intelligence (CTI) is increasingly consumed not by human analysts but by LLM agents that compose multi-step investigations at query time. The harness side of this shift has matured rapidly (planning loops, tool protocols, context management), but the corpus side has not: threat reports and vulnerability databases are still packaged for retrieval-augmented generation, as opaque chunks… ▽ More

    Submitted 19 August, 2026; originally announced August 2026.

    Comments: Preprint

  3. arXiv:2608.16069  [pdf, ps, other

    math.CO

    On the saturation number of the kite graph

    Authors: Huanying Bian, Qing Cui, Shengjin Ji, Fufong Ma

    Abstract: For a fixed graph $H$, a graph $G$ is $H$-saturated if $G$ does not contain a copy of $H$, but adding any edge $e \in E(\overline{G})$ to $G$ creates a copy of $H$. The saturation number $\mathrm{sat}(n,H)$ is the minimum number of edges in an $H$-saturated graph on $n$ vertices. Let $K$ be the kite graph, formed by removing one edge from $ K_4$ and then attaching a pendant edge to a vertex of deg… ▽ More

    Submitted 17 August, 2026; originally announced August 2026.

    Comments: 18 pages, 9 figures

  4. arXiv:2608.02971  [pdf, ps, other

    cs.CL

    Mapping the City Through the Lens of Language Models

    Authors: Wanqi Liu, Rong Zhao, Zhizhou Sha, Qinyu Cui, Yecheng Zhang

    Abstract: Language models often complete an underspecified reference to a city with unstated assumptions about urban size, form, infrastructure, environment, and function. We measure those assumptions without naming places. Ten open-weight checkpoints rate anonymized profiles derived from real morphological urban centres across 40 audited indicators and seven domains. The design combines constrained probabi… ▽ More

    Submitted 7 August, 2026; v1 submitted 3 August, 2026; originally announced August 2026.

  5. arXiv:2607.22777  [pdf

    cs.LG cs.AI

    LC-SEPLM: long-range contact-supervised adaptation for sequence-only protein representation learning

    Authors: Chen Wang, Boming Kang, Qinghua Cui

    Abstract: Protein language models learn transferable sequence representations. However, because they primarily model contextual dependencies along amino-acid sequences, their training objectives do not explicitly constrain the model to learn three-dimensional residue contacts formed after folding . Here, we introduce LC-SEPLM (Long-range Contact-supervised ESM Protein Language Model), which adapts ESM2 with… ▽ More

    Submitted 28 July, 2026; v1 submitted 24 July, 2026; originally announced July 2026.

  6. arXiv:2607.10827  [pdf, ps, other

    math.DG

    A curvature characterization of the Cartan minimal hypersurface in $\mathbb S^5$

    Authors: Qing Cui

    Abstract: Lawson showed that a non-totally geodesic Einstein minimal hypersurface in $\mathbb S^5$ is congruent to the Clifford hypersurface $\mathbb S^2(1/\sqrt2)\times \mathbb S^2(1/\sqrt2).$ It is also known, by work of Cartan and Ôtsuki, that a non-totally geodesic locally conformally flat minimal hypersurface in $\mathbb S^5$ is of Ôtsuki type, including the Clifford hypersurface… ▽ More

    Submitted 12 July, 2026; originally announced July 2026.

    Comments: 22 pages, no figure. Comments are welcome

  7. arXiv:2606.20670  [pdf, ps, other

    cs.LG cs.AI cs.IT

    Towards CSI-Native Foundation Models: A Channel-Adaptive Roadmap for 6G

    Authors: Chenyu Zhang, Xinchen Lyu, Chenshan Ren, Shuhan Liu, Qimei Cui

    Abstract: Wireless foundation models offer a path toward reusable channel state information (CSI) intelligence for sixth-generation (6G) systems. However, existing generic-backbone adaptation and CSI pretraining methods often treat CSI as task tensors rather than propagation-conditioned channel responses, thereby failing to capture the intrinsic time-frequency-spatial geometry of wireless environments. This… ▽ More

    Submitted 12 June, 2026; originally announced June 2026.

    Comments: 7 pages, 5 figures, submmited to IEEE WCM

  8. arXiv:2606.15079  [pdf, ps, other

    cs.CL cs.AI

    Ling and Ring 2.6 Technical Report: Efficient and Instant Agentic Intelligence at Trillion-Parameter Scale

    Authors: Ang Li, Ben Liu, Bin Han, Bin Hu, Bin Jing, Binbin Hu, Bing Li, Cai Chen, Caizhi Tang, Changxin Tian, Chao Huang, Chao Zhang, Chen Liang, Chen Qian, Chengfu Tang, Chengyao Wen, Chilin Fu, Chunwei Wu, Cong Zhang, Cunyin Peng, Daixin Wang, Dalong Zhang, Deng Zhao, Dingnan Jin, Dingyuan Zhu , et al. (193 additional authors not shown)

    Abstract: Efficient and scalable agentic intelligence requires models that can deliver both low-latency responses and strong reasoning capabilities while remaining practical to train, serve, and deploy. In this report, we present Ling-2.6 and Ring-2.6, a family of models designed to address this challenge at scale. Ling-2.6 is optimized for instant response generation and high capability per output token, w… ▽ More

    Submitted 12 June, 2026; originally announced June 2026.

  9. arXiv:2606.11643  [pdf, ps, other

    cs.CL

    Improving Cross-Format Robustness in Language Models with Multi-Format Training

    Authors: June M. Liu, Shaomian Zheng, He Cao, Dingnan Jin, Qing Cui, Jun Zhou

    Abstract: Large language models often remain sensitive to answer format: a question solved correctly in one form may fail in another semantically equivalent form. To study this gap, we define cross-format robustness as the extent to which a model answers the same underlying question consistently across formats. We then compare full-format training with FormatMix, which expands only a subset of training item… ▽ More

    Submitted 10 June, 2026; originally announced June 2026.

  10. arXiv:2606.07378  [pdf, ps, other

    quant-ph cond-mat.mes-hall

    Ferroelectrical Switching as a Probe of Quantum Damping in Magnetic Spin Systems

    Authors: Yuefei Liu, Anna Delin, Olle Eriksson, Erik Sjöqvist, Kaiyou Wang, Qirui Cui

    Abstract: While damped spin dynamics is important for the understanding of magnetic materials, clear signatures of \emph{quantum corrections} to the Gilbert damping mechanism remain elusive. We propose a route to distinguish quantum and classical Gilbert spin damping using ferroelectric control of a magnetic dimer. Ab initio calculations for dimers on ferroelectric substrates show that polarization reversal… ▽ More

    Submitted 5 June, 2026; originally announced June 2026.

    Comments: 7 pages, 3 figures

  11. arXiv:2606.02647  [pdf, ps, other

    math.DG

    A strict upper volume bound for minimal graphs in the unit ball

    Authors: Qing Cui

    Abstract: Let $u$ be a solution of the minimal surface equation on a domain containing the closed unit ball $\overline{B^n}\subset\mathbb R^n$. A classical calibration argument gives $|Graph_u\cap B^{n+1}| \leq \frac12 |\mathbb S^n|.$ A basic question is whether this half-sphere bound is sharp for minimal graphs. We show that it is not. More precisely, for every $n\geq2$, there exists an explicit constant… ▽ More

    Submitted 9 August, 2026; v1 submitted 31 May, 2026; originally announced June 2026.

    Comments: 11 pages, no figure. This version replaces the withdrawn previous version. The previous sharpness claim was incorrect. The present version proves instead a uniform positive gap below the half-sphere bound. All comments are welcome!

  12. arXiv:2606.00409  [pdf, ps, other

    cond-mat.mtrl-sci physics.app-ph

    Data-Driven Discovery of Unconventional Antiferromagnets

    Authors: Qirui Cui, Chenxu Liu, Anna Delin, Kaiyou Wang

    Abstract: Unconventional antiferromagnets combine zero net magnetization with spin-split electronic bands, offering a distinct, important platform for spintronics. Their discovery, however, has so far depended largely on case-by-case studies and on a limited number of compounds with experimentally resolved magnetic structures. Here, we overcome these bottlenecks by resolving magnetic ground states across a… ▽ More

    Submitted 29 May, 2026; originally announced June 2026.

    Comments: 22 pages, 3 figures

  13. arXiv:2605.30723  [pdf, ps, other

    cs.CL

    Skill is Not One-Size-Fits-All: Model-Aware Skill Alignment for LLM Agents

    Authors: Jianxiang Yu, Jiapeng Zhu, Bochen Lin, Qier Cui, Zichen Ding, Xiang Li

    Abstract: LLM agents increasingly retrieve externally curated skills-procedural instructions retrieved at decision time-to improve performance on long-horizon interactive tasks. Existing skill libraries are typically treated as model-agnostic, reusing the same skill formulations across backbones with substantially different capacities and behaviors. However, our controlled experiments across multiple model… ▽ More

    Submitted 28 May, 2026; originally announced May 2026.

  14. arXiv:2605.25836  [pdf, ps, other

    cs.CR cs.AI cs.CL

    TTPrint: Evidence-Grounded TTP Extraction via Diverge-then-Converge Verification

    Authors: Yutong Cheng, Changze Li, Raihan Sultan Pasha Basuki, Qian Cui, Wei Ding, Peng Gao

    Abstract: Extracting MITRE ATT&CK techniques from cyber threat intelligence (CTI) reports is an open-set, multi-label problem requiring both high recall (not missing techniques) and high precision (not hallucinating unsupported ones). Existing methods--rule-based, supervised, and LLM-based--struggle to achieve both: rule-based and supervised approaches lack generalizability across diverse attack description… ▽ More

    Submitted 25 May, 2026; originally announced May 2026.

    Comments: Preprint

  15. arXiv:2605.19762  [pdf, ps, other

    cs.AI cs.CL

    What Really Improves Mathematical Reasoning: Structured Reasoning Signals Beyond Pure Code

    Authors: Yuze Zhao, Junpeng Fang, Lu Yu, Zhenya Huang, Kai Zhang, Qing Cui, Qi Liu, Jun Zhou, Enhong Chen

    Abstract: Code has become a standard component of modern foundation language model (LM) training, yet its role beyond programming remains unclear. We revisit the claim that code improves reasoning through controlled pretraining experiments on a 10T-token corpus with fine-grained domain separation. Our findings are threefold. First, when code is restricted to standalone executable programs and Code-NL data a… ▽ More

    Submitted 19 May, 2026; originally announced May 2026.

    Comments: Accepted by ICML 2026, 22 pages, 10 figures

  16. arXiv:2605.12586  [pdf, ps, other

    cs.CV cs.AI cs.DB

    3D Primitives are a Spatial Language for VLMs

    Authors: Junze Liu, Kun Qian, Florian Dubost, Kai Zhong, Arvind Srinivasan, Nan Chen, Anping Wang, Sam Zhang, Alejandro Mottini, Qingjun Cui, Tian Wang

    Abstract: Vision-language models (VLMs) exhibit a striking paradox: they can generate executable code that reconstructs a 3D scene from geometric primitives with correct object counts, classes, and approximate positions, yet the same models fail at simpler spatial questions on the same image. We show that 3D geometric primitives (cubes, spheres, cylinders, expressed in executable code) serve as a powerful i… ▽ More

    Submitted 12 May, 2026; originally announced May 2026.

  17. arXiv:2605.11601  [pdf, ps, other

    cs.CL cs.AI

    DiffScore: Text Evaluation Beyond Autoregressive Likelihood

    Authors: Wen Lai, Yingli Shen, Dingnan Jin, Qing Cui, Jun Zhou, Maosong Sun, Alexander Fraser

    Abstract: Autoregressive language models are widely used for text evaluation, however, their left-to-right factorization introduces positional bias, i.e., early tokens are scored with only leftward context, conflating architectural asymmetry with true text quality. We propose masked reconstruction as an alternative paradigm, where every token is scored using full bidirectional context. We introduce DiffScor… ▽ More

    Submitted 12 May, 2026; originally announced May 2026.

  18. arXiv:2605.01365  [pdf, ps, other

    cs.CV cs.RO

    VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection

    Authors: Haowen Sun, Shaolong Zhang, Mingyang Li, Chengzhong Ma, Xinzhe Chen, Qiongjie Cui, Xingyu Chen, Zeyang Liu, Xuguang Lan

    Abstract: Open-vocabulary 3D affordance detection requires localizing interaction regions on point clouds given novel affordance descriptions. Recent methods extend multimodal large language models (MLLMs) with special output tokens that are decoded into segmentation masks. However, these tokens are produced through autoregressive generation, which models sequential dependencies rather than spatial neighbor… ▽ More

    Submitted 2 May, 2026; originally announced May 2026.

  19. arXiv:2605.00968  [pdf, ps, other

    eess.SP cs.AI

    Adaptive 3D-RoPE: Physics-Aligned Rotary Positional Encoding for Wireless Foundation Models

    Authors: Chenyu Zhang, Xinchen Lyu, Chenshan Ren, Shuhan Liu, Qimei Cui

    Abstract: Positional encoding plays a pivotal role in determin?ing the extrapolation and generalization performance of wireless foundation models for channel state information (CSI) modeling, latent characterization, and task-specific prediction. However, existing CSI models inherit static or one-dimensional positional priors from natural language and vision architectures, which fundamentally misalign with… ▽ More

    Submitted 1 May, 2026; originally announced May 2026.

    Comments: 13 pages, 7 figures

  20. arXiv:2604.14148  [pdf, ps, other

    cs.CV

    Seedance 2.0: Advancing Video Generation for World Complexity

    Authors: Team Seedance, De Chen, Liyang Chen, Xin Chen, Ying Chen, Zhuo Chen, Zhuowei Chen, Feng Cheng, Tianheng Cheng, Yufeng Cheng, Mojie Chi, Xuyan Chi, Jian Cong, Qinpeng Cui, Fei Ding, Qide Dong, Yujiao Du, Haojie Duanmu, Junliang Fan, Jiarui Fang, Jing Fang, Zetao Fang, Chengjian Feng, Yu Gao, Diandian Gu , et al. (146 additional authors not shown)

    Abstract: Seedance 2.0 is a new native multi-modal audio-video generation model, officially released in China in early February 2026. Compared with its predecessors, Seedance 1.0 and 1.5 Pro, Seedance 2.0 adopts a unified, highly efficient, and large-scale architecture for multi-modal audio-video joint generation. This allows it to support four input modalities: text, image, audio, and video, by integrating… ▽ More

    Submitted 15 April, 2026; originally announced April 2026.

    Comments: Seedance 2.0 Model Card

  21. arXiv:2604.11674  [pdf, ps, other

    cs.RO cs.AI

    AffordSim: A Scalable Data Generator and Benchmark for Affordance-Aware Robotic Manipulation

    Authors: Mingyang Li, Haofan Xu, Haowen Sun, Xinzhe Chen, Sihua Ren, Liqi Huang, Xinyang Sui, Chenyang Miao, Jiawei Ye, Qiongjie Cui, Zeyang Liu, Xingyu Chen, Xuguang Lan

    Abstract: Many everyday robot manipulation skills are affordance-dependent, with success determined by whether the robot contacts the functional object region required by the subsequent action. Current simulation data generators obtain contacts from generic grasp estimators or per-object manual contact annotations, but generic estimators rank stable grasps without task semantics and often select contacts th… ▽ More

    Submitted 11 May, 2026; v1 submitted 13 April, 2026; originally announced April 2026.

  22. arXiv:2603.23967  [pdf, ps, other

    cs.LG

    Wireless communication empowers online scheduling of partially-observable transportation multi-robot systems in a smart factory

    Authors: Yaxin Liao, Qimei Cui, Kwang-Cheng Chen, Xiong Li, Jinlian Chen, Xiyu Zhao, Xiaofeng Tao, Ping Zhang

    Abstract: Achieving agile and reconfigurable production flows in smart factories depends on online multi-robot task assignment (MRTA), which requires online collision-free and congestion-free route scheduling of transportation multi-robot systems (T-MRS), e.g., collaborative automatic guided vehicles (AGVs). Due to the real-time operational requirements and dynamic interactions between T-MRS and production… ▽ More

    Submitted 25 March, 2026; originally announced March 2026.

  23. arXiv:2603.22830  [pdf

    cond-mat.mtrl-sci physics.comp-ph

    Profound impacts of interlayer interactions in bilayer altermagnetic V2S2O

    Authors: Siqi Xu, Qilong Cui, Shaowen Xu, Xianbo Chenwei, Jiahao Zhang, Ruixue Li, Yuan Li, Gaofeng Xu, Fanhao Jia

    Abstract: Two-dimensional altermagnets exhibit exceptional potential for low-power spintronics via nonrelativistic spin splitting and zero net magnetization. Here, we systematically investigate the influence of interlayer interactions on the electronic, magnetic and quantum transport properties of bilayer vanadium oxysulfide (V2S2O), a prototypical layered altermagnet, using DFT and NEGF calculations. Our r… ▽ More

    Submitted 24 March, 2026; originally announced March 2026.

  24. arXiv:2603.22450  [pdf, ps, other

    cs.CV cs.GR

    Static Scene Reconstruction from Dynamic Egocentric Videos

    Authors: Qifei Cui, Patrick Chen

    Abstract: Egocentric videos present unique challenges for 3D reconstruction due to rapid camera motion and frequent dynamic interactions. State-of-the-art static reconstruction systems, such as MapAnything, often degrade in these settings, suffering from catastrophic trajectory drift and "ghost" geometry caused by moving hands. We bridge this gap by proposing a robust pipeline that adapts static reconstruct… ▽ More

    Submitted 23 March, 2026; originally announced March 2026.

  25. arXiv:2603.18075  [pdf, ps, other

    astro-ph.HE gr-qc

    Waveforms and Fluxes of Generic Extreme-Mass-Ratio Inspirals with a Spinning Secondary

    Authors: Qiuxin Cui, Wen-Biao Han

    Abstract: Extreme mass-ratio inspirals (EMRIs), comprising a stellar-mass compact object (CO) orbiting a supermassive black hole (BH), are key targets for future space-based gravitational-wave (GW) observatories. Incorporating the spin of the secondary body into waveform models not only enhances measurement precision but also offers insight into the spin distribution of stellar-mass COs. In this work, we co… ▽ More

    Submitted 14 July, 2026; v1 submitted 18 March, 2026; originally announced March 2026.

  26. Star formation in the circumgalactic high-velocity cloud Complex H

    Authors: Zhihong He, Wenkang Pang, Kun Wang, Yangping Luo, Qian Cui

    Abstract: The accretion of metal-poor gas sustains galactic star formation. In the Milky Way, this process is fueled by high-velocity clouds (HVCs), yet their fundamental properties have remained elusive in the absence of stellar tracers. Here we report a binary open cluster within HVC Complex H. With an age of 11.2 +- 0.6 Myr and a subsolar metallicity of 0.05(+0.05-0.02) Zsun, the clusters provide a direc… ▽ More

    Submitted 11 March, 2026; originally announced March 2026.

    Comments: 31 pages, 10 figures, published on Nature Astronomy (https://www.nature.com/articles/s41550-026-02814-9)

  27. arXiv:2603.04475  [pdf, ps, other

    physics.ins-det nucl-ex

    Commissioning and Full Realization of the PLASEN System at BRIF

    Authors: W. C. Mei, H. R. Hu, Y. F. Guo, Z. Yan, X. F. Yang, S. J. Chen, D. Y. Chen, Y. P. Lin, Y. S. Liu, C. Zhang, Y. P. Jing, T. X. Gao, X. Shen, Y. Y. Jia, Y. T. Lin, H. X. Zhang, S. W. Bai, B. Tang, X. Ma, G. F. Song, S. Ye, M. Y. Lu, J. Y. Dong, B. K. Dong, J. H. Lv , et al. (15 additional authors not shown)

    Abstract: A PLASEN (Precision LAser Spectroscopy for Exotic Nuclei) system, consisting of a compact radio-frequency quadrupole cooler-buncher (RFQ-cb) and a collinear resonance ionization spectroscopy setup, has now been fully commissioned with radioactive ion beams at the Beijing Radioactive Ion-beam Facility (BRIF). Using both stable and radioactive Rb ion beams from BRIF, we demonstrated that the large b… ▽ More

    Submitted 4 March, 2026; originally announced March 2026.

    Journal ref: Phys. Rev. Research 8, 023286 (2026)

  28. arXiv:2603.00031  [pdf, ps, other

    cs.CL cs.LG

    GRIP: Geometric Refinement and Adaptive Information Potential for Data Efficiency

    Authors: Changhao Wang, Jiaolong Yang, Xinhao Yao, Yunfei Yu, Peng Jiao, Lu Yu, Junpeng Fang, Riccardo Cantoro, Qing Cui, Jun Zhou

    Abstract: The performance of Large Language Models (LLMs) is increasingly governed by data efficiency rather than raw scaling volume. However, existing selection methods often decouple global distribution balancing from local instance selection, compromising the hierarchical integrity of the training set. We introduce \textbf{GRIP} (Geometric Refinement and Adaptive Information Potential), a framework that… ▽ More

    Submitted 4 February, 2026; originally announced March 2026.

  29. arXiv:2602.21104  [pdf, ps, other

    cs.LG cs.DS

    Ski Rental with Distributional Predictions of Unknown Quality

    Authors: Qiming Cui, Michael Dinitz

    Abstract: We revisit the central online problem of ski rental in the "algorithms with predictions" framework from the point of view of distributional predictions. Ski rental was one of the first problems to be studied with predictions, where a natural prediction is simply the number of ski days. But it is both more natural and potentially more powerful to think of a prediction as a distribution p-hat over t… ▽ More

    Submitted 24 February, 2026; originally announced February 2026.

  30. arXiv:2602.13773  [pdf, ps, other

    cs.LG

    On Representation Redundancy in Large-Scale Instruction Tuning Data Selection

    Authors: Youwei Shu, Shaomian Zheng, Dingnan Jin, Wenjie Qu, Ziyao Guo, Qing Cui, Jun Zhou, Jiaheng Zhang

    Abstract: Data quality is a crucial factor in large language models training. While prior work has shown that models trained on smaller, high-quality datasets can outperform those trained on much larger but noisy or low-quality corpora, systematic methods for industrial-scale data selection in instruction tuning remain underexplored. In this work, we study instruction-tuning data selection through the lens… ▽ More

    Submitted 14 February, 2026; originally announced February 2026.

  31. arXiv:2602.05411  [pdf, ps, other

    astro-ph.GA

    Mergers Drive Structural Complexity but Not Starbursts in Lyman-$α$ Emitters at $3 < z < 4$: A JWST Spatially Resolved View

    Authors: Qi Song, F. S. Liu, Jian Ren, Tianfu Gao, Pinsong Zhao, Qifan Cui, Yubin Li, Hao Mo, Guanghuan Wang

    Abstract: Recent observations with the James Webb Space Telescope (JWST) reveal that the merger fraction among Ly$α$ emitters (LAEs) at redshifts $z > 3$ is significantly higher than previously estimated. In this study, we focus on three high signal-to-noise merging LAE systems at $3 < z < 4$, selected from the VLT/MUSE-Deep survey in the GOODS-S field. We combine new \textit{JWST}/NIRCam broadband and medi… ▽ More

    Submitted 5 February, 2026; originally announced February 2026.

    Comments: Accepted for publication in RAA

  32. arXiv:2602.03772  [pdf, ps, other

    cs.LG cs.AI

    UniGeM: Unifying Data Mixing and Selection via Geometric Exploration and Mining

    Authors: Changhao Wang, Yunfei Yu, Xinhao Yao, Jiaolong Yang, Riccardo Cantoro, Chaobo Li, Qing Cui, Jun Zhou

    Abstract: The scaling of Large Language Models (LLMs) is increasingly limited by data quality. Most methods handle data mixing and sample selection separately, which can break the structure in code corpora. We introduce \textbf{UniGeM}, a framework that unifies mixing and selection by treating data curation as a \textit{manifold approximation} problem without training proxy models or relying on external ref… ▽ More

    Submitted 3 February, 2026; originally announced February 2026.

  33. arXiv:2601.20178  [pdf, ps, other

    eess.SP

    Coverage Performance Analysis of FAS-enhanced LoRa Wide Area Networks under both Co-SF and Inter-SF Interference

    Authors: Gaoze Mu, Yanzhao Hou, Mingjie Chen, Yuanyu Hu, Yongan Zheng, Qimei Cui, Xiaofeng Tao

    Abstract: This paper presents an analytical framework for evaluating the coverage performance of the fluid antenna system (FAS)-enhanced LoRa wide-area networks (LoRaWANs). We investigate the effects of large-scale pathloss in LoRaWAN, small-scale fading characterized by FAS, and dense interference (i.e., packet collisions under the ALOHA protocol) arising from randomly deployed end devices (EDs). Both co-s… ▽ More

    Submitted 19 April, 2026; v1 submitted 27 January, 2026; originally announced January 2026.

    Comments: 6 pages, 3 figures

  34. arXiv:2601.18200  [pdf, ps, other

    cs.LG cs.AI

    HeterCSI: Channel-Adaptive Heterogeneous CSI Pretraining Framework for Generalized Wireless Foundation Models

    Authors: Chenyu Zhang, Xinchen Lyu, Chenshan Ren, Shuhan Liu, Qimei Cui, Xiaofeng Tao

    Abstract: Wireless foundation models promise transformative capabilities for channel state information (CSI) processing across diverse 6G network applications, yet face fundamental challenges due to the inherent dual heterogeneity of CSI across both scale and scenario dimensions. However, current pretraining approaches either constrain inputs to fixed dimensions or isolate training by scale, limiting the ge… ▽ More

    Submitted 26 January, 2026; originally announced January 2026.

    Comments: 13 pages, 8 figures

  35. arXiv:2601.09017  [pdf

    cs.CL

    Multicultural Spyfall: Assessing LLMs through Dynamic Multilingual Social Deduction Game

    Authors: Haryo Akbarianto Wibowo, Alaa Elsetohy, Qinrong Cui, Alham Fikri Aji

    Abstract: The rapid advancement of Large Language Models (LLMs) has necessitated more robust evaluation methods that go beyond static benchmarks, which are increasingly prone to data saturation and leakage. In this paper, we propose a dynamic benchmarking framework for evaluating multilingual and multicultural capabilities through the social deduction game Spyfall. In our setup, models must engage in strate… ▽ More

    Submitted 13 January, 2026; originally announced January 2026.

    MSC Class: 68T50

  36. arXiv:2512.20640  [pdf, ps, other

    cs.NI cs.MA

    Reflection-Driven Self-Optimization 6G Agentic AI RAN via Simulation-in-the-Loop Workflows

    Authors: Yunhao Hu, Xinchen Lyu, Chenshan Ren, Keda Chen, Qimei Cui, Xiaofeng Tao

    Abstract: The escalating complexity of sixth-generation (6G) networks demands unprecedented levels of autonomy beyond the capabilities of traditional optimization-based and current AI-based resource management approaches. While agentic AI has emerged as a promising paradigm for autonomous RAN, current frameworks provide sophisticated reasoning capabilities but lack mechanisms for empirical validation and se… ▽ More

    Submitted 21 April, 2026; v1 submitted 8 December, 2025; originally announced December 2025.

  37. arXiv:2512.13507  [pdf, ps, other

    cs.CV

    Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model

    Authors: Team Seedance, Heyi Chen, Siyan Chen, Xin Chen, Yanfei Chen, Ying Chen, Zhuo Chen, Feng Cheng, Tianheng Cheng, Xinqi Cheng, Xuyan Chi, Jian Cong, Jing Cui, Qinpeng Cui, Qide Dong, Junliang Fan, Jing Fang, Zetao Fang, Chengjian Feng, Han Feng, Mingyuan Gao, Yu Gao, Dong Guo, Qiushan Guo, Boyang Hao , et al. (172 additional authors not shown)

    Abstract: Recent strides in video generation have paved the way for unified audio-visual generation. In this work, we present Seedance 1.5 pro, a foundational model engineered specifically for native, joint audio-video generation. Leveraging a dual-branch Diffusion Transformer architecture, the model integrates a cross-modal joint module with a specialized multi-stage data pipeline, achieving exceptional au… ▽ More

    Submitted 23 December, 2025; v1 submitted 15 December, 2025; originally announced December 2025.

    Comments: Seedance 1.5 pro Technical Report

  38. arXiv:2512.09485  [pdf, ps, other

    cs.CR cs.AI

    Advancing LLM-Based Security Automation with Customized Group Relative Policy Optimization for Zero-Touch Networks

    Authors: Xinye Cao, Yihan Lin, Guoshun Nan, Qinchuan Zhou, Yuhang Luo, Yurui Gao, Zeliang Zhang, Haolang Lu, Qimei Cui, Yanzhao Hou, Xiaofeng Tao, Tony Q. S. Quek

    Abstract: Zero-Touch Networks (ZTNs) represent a transformative paradigm toward fully automated and intelligent network management, providing the scalability and adaptability required for the complexity of sixth-generation (6G) networks. However, the distributed architecture, high openness, and deep heterogeneity of 6G networks expand the attack surface and pose unprecedented security challenges. To address… ▽ More

    Submitted 10 December, 2025; originally announced December 2025.

    Comments: Accepted by IEEE JSAC. This work has been submitted to the IEEE for possible publication

  39. arXiv:2511.21707  [pdf, ps, other

    cs.NI cs.AI

    Sensing and Understanding the World over Air: A Large Multimodal Model for Mobile Networks

    Authors: Zhuoran Duan, Yuhao Wei, Guoshun Nan, Zijun Wang, Yan Yan, Lihua Xiong, Yuhan Ran, Ji Zhang, Jian Li, Qimei Cui, Xiaofeng Tao, Tony Q. S. Quek

    Abstract: Large models (LMs), such as ChatGPT, have made a significant impact across diverse domains and hold great potential to facilitate the evolution of network intelligence. Wireless-native multi-modal large models (WMLMs) can sense and understand the physical world through multi-modal data, serving as a key enabler that integrates communication, sensing, and intelligence, and thus they can boost vario… ▽ More

    Submitted 17 November, 2025; originally announced November 2025.

  40. arXiv:2511.12865  [pdf, ps, other

    cs.LG cs.AI

    An approach of deep reinforcement learning for maximizing the net present value of stochastic projects

    Authors: Wei Xu, Fan Yang, Qinyuan Cui, Zhi Chen

    Abstract: This paper investigates a project with stochastic activity durations and cash flows under discrete scenarios, where activities must satisfy precedence constraints generating cash inflows and outflows. The objective is to maximize expected net present value (NPV) by accelerating inflows and deferring outflows. We formulate the problem as a discrete-time Markov Decision Process (MDP) and propose a D… ▽ More

    Submitted 16 November, 2025; originally announced November 2025.

  41. arXiv:2511.11990  [pdf, ps, other

    cs.AI

    Improving Autoformalization Using Direct Dependency Retrieval

    Authors: Shaoqi Wang, Lu Yu, Siwei Lou, Feng Yan, Chunjie Yang, Qing Cui, Jun Zhou

    Abstract: The convergence of deep learning and formal mathematics has spurred research in formal verification. Statement autoformalization, a crucial first step in this process, aims to translate informal descriptions into machine-verifiable representations but remains a significant challenge. The core difficulty lies in the fact that existing methods often suffer from a lack of contextual awareness, leadin… ▽ More

    Submitted 1 January, 2026; v1 submitted 14 November, 2025; originally announced November 2025.

  42. arXiv:2510.22115  [pdf, ps, other

    cs.CL cs.AI

    Every Activation Boosted: Scaling General Reasoner to 1 Trillion Open Language Foundation

    Authors: Ling Team, Ang Li, Ben Liu, Binbin Hu, Bing Li, Bingwei Zeng, Borui Ye, Caizhi Tang, Changxin Tian, Chao Huang, Chao Zhang, Chen Qian, Chenchen Ju, Chenchen Li, Chengfu Tang, Chilin Fu, Chunshao Ren, Chunwei Wu, Cong Zhang, Cunyin Peng, Dafeng Xu, Daixin Wang, Dalong Zhang, Dingnan Jin, Dingyuan Zhu , et al. (117 additional authors not shown)

    Abstract: We introduce Ling 2.0, a series reasoning-oriented language foundation built upon the principle that every activation boosts reasoning capability. Designed to scale from tens of billions to one trillion parameters under a unified Mixture-of-Experts (MoE) paradigm, Ling 2.0 emphasizes high sparsity, cross-scale consistency, and efficiency guided by empirical scaling laws. The series includes three… ▽ More

    Submitted 6 November, 2025; v1 submitted 24 October, 2025; originally announced October 2025.

    Comments: Ling 2.0 Technical Report

  43. arXiv:2510.19245  [pdf, ps, other

    cs.CY cs.AI cs.HC cs.LG cs.MM

    See, Think, Act: Online Shopper Behavior Simulation with VLM Agents

    Authors: Yimeng Zhang, Jiri Gesi, Ran Xue, Tian Wang, Ziyi Wang, Yuxuan Lu, Sinong Zhan, Huimin Zeng, Qingjun Cui, Yufan Guo, Jing Huang, Mubarak Shah, Dakuo Wang

    Abstract: LLMs have recently demonstrated strong potential in simulating online shopper behavior. Prior work has improved action prediction by applying SFT on action traces with LLM-generated rationales, and by leveraging RL to further enhance reasoning capabilities. Despite these advances, current approaches rely on text-based inputs and overlook the essential role of visual perception in shaping human dec… ▽ More

    Submitted 22 October, 2025; originally announced October 2025.

  44. arXiv:2510.16418  [pdf, ps, other

    cs.DC

    FourierCompress: Layer-Aware Spectral Activation Compression for Efficient and Accurate Collaborative LLM Inference

    Authors: Jian Ma, Xinchen Lyu, Jun Jiang, Longhao Zou, Chenshan Ren, Qimei Cui, Xiaofeng Tao

    Abstract: Collaborative large language model (LLM) inference enables real-time, privacy-preserving AI services on resource-constrained edge devices by partitioning computational workloads between client devices and edge servers. However, this paradigm is severely hindered by communication bottlenecks caused by the transmission of high-dimensional intermediate activations, exacerbated by the autoregressive d… ▽ More

    Submitted 18 October, 2025; originally announced October 2025.

  45. arXiv:2510.10458  [pdf, ps, other

    math.CO

    Some results on minimum saturated graphs

    Authors: Chenke Zhang, Qing Cui, Jinze Hu, Erfei Yue, Shengjin Ji

    Abstract: Let $G$ be a graph and $\mathcal{F}$ be a family of graphs. We say a graph $G$ is $\mathcal{F}$-saturated if $G$ does not contain any member in $\mathcal{F}$ and for any $e\in E(\overline{G})$, $G+e$ creates a copy of some member in $ \mathcal{F}$. The saturation number of $\mathcal{F}$ is the minimum number of edges of an $\mathcal{F}$-saturated graphs with $n$ vertices, denoted by… ▽ More

    Submitted 12 October, 2025; originally announced October 2025.

    Comments: 16 pages,5 figures

  46. arXiv:2510.08390  [pdf, ps, other

    astro-ph.SR

    Gaia DR3 Open Cluster Cepheids: A Unified Catalog with Calibrated Period-Age and Period-Wesenheit Relations

    Authors: Shunhong Deng, Zhihong He, Anbing Ren, Qian Cui, Xiaoyue Zhou, Liming Peng, Chenxin Wang, Ziang Chen, Yangping Luo, Kun Wang

    Abstract: Classical Cepheids (CCs) in Galactic open clusters (OCs) provide essential observational constraints for calibrating the period-age relation (PAR) and the period-Wesenheit relation (PWR) of CCs. However, distant and long-period OC Cepheids remain limited, while the confirmed samples still require more precise determinations of their physical properties, such as ages and extinctions. In this work,… ▽ More

    Submitted 31 October, 2025; v1 submitted 9 October, 2025; originally announced October 2025.

    Comments: 29 pages, 6 figures, 3 tables, 6 appendix figures, 3 appendix tables, Accepted for publication in AJ, minor corrections for typo

  47. arXiv:2510.04664  [pdf, ps, other

    cs.IT

    Learning Function-to-Function Mappings: A Fourier Neural Operator for Next-Generation MIMO Systems

    Authors: Jian Xiao, Ji Wang, Qi Sun, Qimei Cui, Xingwang Li, Dusit Niyato, Chih-Lin I

    Abstract: Next-generation multiple-input multiple-output (MIMO) systems, characterized by extremely large-scale arrays, holographic surfaces, three-dimensional architectures, and flexible antennas, are poised to deliver unprecedented data rates, spectral efficiency and stability. However, these advancements introduce significant challenges for physical layer signal processing, stemming from complex near-fie… ▽ More

    Submitted 6 October, 2025; originally announced October 2025.

  48. arXiv:2510.04028  [pdf, ps, other

    cs.LG cs.AI

    The Debate on RLVR Reasoning Capability Boundary: Shrinkage, Expansion, or Both? A Two-Stage Dynamic View

    Authors: Xinhao Yao, Lu Yu, Xiaolin Hu, Fengwei Teng, Qing Cui, Jun Zhou, Yong Liu

    Abstract: The ongoing debate on whether reinforcement learning with verifiable rewards (RLVR) expands or shrinks the reasoning capabilities of large language models (LLMs) remains unresolved. Some studies contend that RLVR mainly improves sampling efficiency but at the expense of diversity and exploratory capacity, resulting in capability boundary shrinkage. In contrast, others demonstrate that prolonged tr… ▽ More

    Submitted 5 October, 2025; originally announced October 2025.

  49. arXiv:2509.25773  [pdf, ps, other

    cs.CV cs.AI cs.CL

    v-HUB: A Benchmark for Video Humor Understanding from Vision and Sound

    Authors: Zhengpeng Shi, Yanpeng Zhao, Jianqun Zhou, Yuxuan Wang, Qinrong Cui, Wei Bi, Songchun Zhu, Bo Zhao, Zilong Zheng

    Abstract: AI models capable of comprehending humor hold real-world promise -- for example, enhancing engagement in human-machine interactions. To gauge and diagnose the capacity of multimodal large language models (MLLMs) for humor understanding, we introduce v-HUB, a novel video humor understanding benchmark. v-HUB comprises a curated collection of non-verbal short videos, reflecting real-world scenarios w… ▽ More

    Submitted 1 June, 2026; v1 submitted 30 September, 2025; originally announced September 2025.

    Comments: 24 pages, 9 figures

  50. arXiv:2509.24062  [pdf, ps, other

    astro-ph.GA

    Identifying Dust-lane Spheroidal Galaxies in DESI Legacy Imaging Surveys Using Semi-Supervised Methods

    Authors: Zhijian Luo, Jianzhen Chen, Wenxiang Pei, Hubing Xiao, Shaohua Zhang, Qifan Cui, Chenggang Shu

    Abstract: Dust-lane spheroidal galaxies (DLSGs) are unique astrophysical systems that exhibit the morphology of early-type galaxies (ETGs) but are distinguished by prominent dust lanes. Recent studies propose that they form through minor mergers between ETGs and gas-rich dwarf galaxies, offering a window into the interstellar medium (ISM) of ETGs and star formation triggered by small-scale interactions. How… ▽ More

    Submitted 28 September, 2025; originally announced September 2025.