Skip to main content
arXiv is now an independent nonprofit! Learn more

Showing 1–50 of 2,558 results for author: Cheng, J

.
  1. arXiv:2608.21006  [pdf, ps, other

    hep-ex

    Evidence for $η_{c}(2S)\to p\bar{p}π^{+}π^{-}π^{0}$ and observation of $χ_{cJ} \to p\bar{p}π^{+}π^{-}π^{0}$

    Authors: BESIII Collaboration, M. Ablikim, M. N. Achasov, P. Adlarson, X. C. Ai, C. S. Akondi, R. Aliberti, A. Amoroso, Q. An, Y. H. An, Y. Bai, O. Bakina, H. R. Bao, X. L. Bao, M. Barbagiovanni, V. Batozskaya, K. Begzsuren, N. Berger, M. Berlowski, M. B. Bertani, D. Bettoni, F. Bianchi, E. Bianco, A. Bortone, I. Boyko , et al. (750 additional authors not shown)

    Abstract: Using $(2.712\pm0.014)\times 10^9$ $ψ(3686)$ events collected by the BESIII detector at the BEPCII collider, the $ψ(3686) \to γp\bar{p}π^+π^-π^0$ process is investigated. Evidence for the decay of $η_{c}(2S)\to p\bar{p}π^{+}π^{-}π^{0}$ is found with a signal significance of 3.3$σ$. The product of branching fractions of… ▽ More

    Submitted 21 August, 2026; originally announced August 2026.

  2. arXiv:2608.19383  [pdf, ps, other

    stat.ME stat.ML

    Causal Generalization of Continuous Treatment Effects under Covariate Shift

    Authors: Jay Jojo Cheng, Guanhua Chen

    Abstract: Average dose-response functions are widely used to summarize causal effects of continuous treatments, but most existing methods assume that the observed sample represents the target population. We study a covariate-shift setting in which covariates, treatment, and outcome are observed in a labelled source sample, while only covariates are observed in the target sample. We develop a two-sample loca… ▽ More

    Submitted 19 August, 2026; originally announced August 2026.

  3. arXiv:2608.18623  [pdf, ps, other

    physics.chem-ph

    UBio-MolFM: Enabling Biomolecular Dynamics at DFT Accuracy and $10^5$ Atoms with One Untuned Potential

    Authors: Lin Huang, Frank Peng, JiaJun Cheng, Zion Wang, Hao Yin, Hao Li, Ji Zhang, Jack Jia, Junping Zhao, Arthur Jiang, Jia Zhang

    Abstract: Ion conduction, membrane permeation and metal recognition hinge on electronic structure, yet first-principles simulation reaches only hundreds of atoms. UBio-MolFM lifts that ceiling: a foundation model trained on 160 million quantum-chemical labels, its receptive field spanning non-covalent distances at near-linear cost. The barrier is cost, not principle. One untuned potential keeps force error… ▽ More

    Submitted 19 August, 2026; originally announced August 2026.

  4. arXiv:2608.16214  [pdf, ps, other

    hep-ex

    First measurements of the branching fractions of $J/ψ$ and $ψ(3686) \to Σ^{0} \barΣ^{0}η$

    Authors: BESIII Collaboration, M. Ablikim, M. N. Achasov, P. Adlarson, X. C. Ai, C. S. Akondi, R. Aliberti, A. Amoroso, Q. An, Y. H. An, Y. Bai, O. Bakina, H. R. Bao, X. L. Bao, M. Barbagiovanni, V. Batozskaya, K. Begzsuren, N. Berger, M. Berlowski, M. B. Bertani, D. Bettoni, F. Bianchi, E. Bianco, A. Bortone, I. Boyko , et al. (750 additional authors not shown)

    Abstract: Based on $(10087 \pm 44) \times 10^6$ $J/ψ$ and $(2712 \pm 14) \times 10^6$ $ψ(3686)$ events collected with the BESIII detector at the BEPCII collider, the hadronic decays $J/ψ\to Σ^{0} \barΣ^{0} η$ and $ψ(3686) \to Σ^{0} \barΣ^{0} η$ are observed for the first time. The corresponding branching fractions are measured to be… ▽ More

    Submitted 17 August, 2026; originally announced August 2026.

  5. arXiv:2608.16076  [pdf, ps, other

    hep-ex

    Measurement of Branching Fraction and Transition Magnetic Moment of the Hyperon Dalitz Decay $Σ^0 \rightarrow Λe^+e^-$

    Authors: BESIII Collaboration, M. Ablikim, M. N. Achasov, P. Adlarson, X. C. Ai, R. Aliberti, A. Amoroso, Q. An, Y. Bai, O. Bakina, Y. Ban, H. -R. Bao, V. Batozskaya, K. Begzsuren, N. Berger, M. Berlowski, M. B. Bertani, D. Bettoni, F. Bianchi, E. Bianco, A. Bortone, I. Boyko, R. A. Briere, A. Brueggemann, H. Cai , et al. (683 additional authors not shown)

    Abstract: Based on a data sample of 10 billion $J/ψ$ events collected with the BESIII detector operating at the BEPCII collider, the Dalitz decay $Σ^0 \rightarrow Λe^+e^-$ is studied experimentally for the first time. The $Σ^0$ hyperons are produced through the process $J/ψ\rightarrow Σ^0\barΣ^0$ and analyzed using a double-tag method. The absolute branching fraction is measured to be… ▽ More

    Submitted 17 August, 2026; originally announced August 2026.

    Comments: 10 pages, 3 figures; supplemental material included

  6. arXiv:2608.15693  [pdf, ps, other

    cs.AI cs.LG

    Large Models for Small Devices: Recent Advances and Empirical Analysis of Edge AI Deployment

    Authors: Subhransu Das, Jiaming Cheng, Arnav Kumar, Sadia Afrose, Mingzhe Han, Michael Silagy, Shreya Palande, Brijesh Soni, Rajiv Ramnath

    Abstract: Running large AI models on resource-constrained edge devices requires model compression to reduce model size and computation. What compresses well, however, need not deploy well. We survey dozens of recent works that report compression results on real hardware and extract practical deployment guidelines from them. Following these guidelines, we deploy compact language and image models on GPU, CPU,… ▽ More

    Submitted 16 August, 2026; originally announced August 2026.

    Comments: Parts of this work were presented at the IEEE Consumer Communications & Networking Conference (CCNC), Las Vegas, NV, USA, January 2026

  7. arXiv:2608.15045  [pdf, ps, other

    cs.CV

    MOSS-VL Technical Report

    Authors: Pengyu Wang, Chenkun Tan, Shaojun Zhou, Qirui Zhou, Yanxin Chen, Xingyang He, Huazheng Zeng, Jijun Cheng, Chenghao Wang, Xiaomeng Qian, Pengfei Wang, Zhan Huang, Shanqing Gao, Wei Huang, Longjun Cao, Wu Ran, Jie Liu, Changtai Zhu, Hongkai Wang, Yixian Tian, Chenghao Liu, Zhen Ye, Xinghao Wang, Botian Jiang, Guoguo Feng , et al. (7 additional authors not shown)

    Abstract: We present MOSS-VL, an open vision-language model family that treats real-time interaction -- perceiving while it speaks -- as a first-class capability. It is co-designed across the stack: the language decoder attends to vision only through gated cross-attention, so the model can naturally see incoming frames while generating; a synthesized interaction corpus supervises when to speak, when to stay… ▽ More

    Submitted 15 August, 2026; originally announced August 2026.

    Comments: 22 pages. Project page: https://openmoss.ai/MOSS-VL/

  8. arXiv:2608.14216  [pdf, ps, other

    cs.CR

    MazeRunner: Nonlinear Task and Clue Orchestration for LLM-driven Black-Box Automated Penetration Testing

    Authors: Zhenyuan Li, Yi Jiang, Junjie Cheng, Yaokun Li, Jing Qiu, Shouling Ji

    Abstract: Penetration testing is essential yet resource-intensive. Although large language models (LLMs) show promise for automating security auditing, existing agents mainly execute end-to-end workflows in simplified linear scenarios. Real-world black-box testing is fundamentally nonlinear: the attack graph is initially unknown and must be incrementally inferred from environmental feedback. Observations ma… ▽ More

    Submitted 14 August, 2026; originally announced August 2026.

  9. Practical Lossless Volumetric Medical Image Compression via Tri-plane Context Tree Learning

    Authors: Yuanchao Bai, Yifan Zhao, Kai Wang, Yuanbo Du, Jie Cheng, Teng Fang, Xianming Liu, Wen Gao

    Abstract: Lossless compression of volumetric medical images is of paramount importance for clinical and research applications where data fidelity is essential. Traditional compression methods are often limited in efficiency due to rigid, handcrafted models. Conversely, deep neural network (DNN)-based compression methods, while effective, demand substantial computational resources, hindering deployment in re… ▽ More

    Submitted 13 August, 2026; originally announced August 2026.

  10. arXiv:2608.12793  [pdf, ps, other

    hep-ex

    High-precision measurement of the space-like $η^\prime$ transition form factor

    Authors: BESIII Collaboration, M. Ablikim, M. N. Achasov, P. Adlarson, X. C. Ai, C. S. Akondi, R. Aliberti, A. Amoroso, Q. An, Y. H. An, M. S. Anderson, Y. Bai, O. Bakina, H. R. Bao, X. L. Bao, M. Barbagiovanni, V. Batozskaya, K. Begzsuren, N. Berger, M. Berlowski, M. B. Bertani, D. Bettoni, F. Bianchi, E. Bianco, A. Bortone , et al. (758 additional authors not shown)

    Abstract: Using a data sample corresponding to an integrated luminosity of $20.3\ \text{fb}^{-1}$, collected with the BESIII detector at a center-of-mass energy of $3.773\ \text{GeV}$ at the BEPCII collider, we report a precision measurement of the product $Q^2|F(Q^2)|$, where $F(Q^2)$ is the single-virtual space-like transition form factor of the $η'$ meson and $Q^2$ is the squared momentum transfer of the… ▽ More

    Submitted 12 August, 2026; originally announced August 2026.

  11. arXiv:2608.12597  [pdf, ps, other

    cs.LG cs.AI

    Predicting When Random Low-Dimensional Reparameterizations Train Neural Networks

    Authors: Andrew Cheng, Ali Eslamian, Jie Cheng, Mehdi Zargham, Qiang Cheng

    Abstract: Neural networks can often be trained or fine-tuned through random low-dimensional reparameterization, where a small latent vector is mapped into a full parameter update by a frozen random map. This raises a practical question: how large must the latent search space be to reach a low-loss region? We first express the known accessibility transition in an equivalent conic form, centered for compact c… ▽ More

    Submitted 12 August, 2026; originally announced August 2026.

  12. arXiv:2608.12564  [pdf, ps, other

    cs.LG

    Scaling Automatic Research Agents via World Models

    Authors: Xiyuan Yang, Sheikh Sarwar, Jingru Cheng, Zhan Shi, Duanshun Li, Huiyuan Chen, Haiyang Zhang, Chenlei Guo, Jingrui He, Zhenyu Liao

    Abstract: Automating empirical research is a long-standing direction of AI. Recent automatic research (AutoResearch) agents bring this goal within reach, as modern LLMs show the capability to independently implement solutions and learn from the execution outcomes. Behind these gains, post-training (especially RL) plays a central role. In this paper, we identify a fundamental tension when scaling RL for thes… ▽ More

    Submitted 12 August, 2026; originally announced August 2026.

  13. arXiv:2608.12009  [pdf, ps, other

    math.OC cs.LG

    Adaptive Bregman Proximal Stochastic Gradient with a Stabilized Barzilai--Borwein Step Size

    Authors: Chenhan Jin, Shengze Xu, Binghui Xie, Kaiwen Zhou, Fan Jia, James Cheng, Tieyong Zeng

    Abstract: Bregman proximal stochastic gradient (BPSG) methods bring variance-reduced composite optimization to objectives whose geometry is poorly captured by Euclidean smoothness. Their performance, however, remains sensitive to the step size: raw stochastic curvature estimates can fluctuate sharply, whereas line searches add repeated proximal evaluations. We introduce Ada-BPSG, a line-search-free BPSG met… ▽ More

    Submitted 12 August, 2026; originally announced August 2026.

  14. arXiv:2608.10915  [pdf, ps, other

    cs.AI

    ComBodied Agents: a New Paradigm of Human-Centric Agentic AI

    Authors: Qianggang Ding, Xingyao Wang, Rui Feng, Zhibin Wang, Feixiang Yao, Kelong Mao, Hao Sun, Zhiyao Luo, Jiankai Tang, Lei Li, Jiadong Guo, Minheng Ni, Weicong Lin, Chenxi Yang, Hongxiang Gao, Zhenghua Chen, Yang Bai, Min Wu, Jun Cheng, Huazhu Fu, Dacheng Tao, Bang Liu

    Abstract: After an older adult misses a medication dose, a software agent can send another reminder and an embodied agent can bring the medication. Yet neither explains whether the person forgot, is confused, has side effects, or deliberately refused, nor what support is appropriate. This reveals a structural gap in Agentic AI: Digital Agents primarily transform software states, while Embodied Agents transf… ▽ More

    Submitted 12 August, 2026; v1 submitted 11 August, 2026; originally announced August 2026.

    Comments: 38 pages, 6 figures, 10 tables

  15. arXiv:2608.08374  [pdf, ps, other

    cs.CV

    Gated Spatial Redundancy Projection for Pathology Transformer Attentions

    Authors: Zhiyuan Yang, Jiahao Cheng, Vincent Quoc-Huy Trinh, Mahdi S. Hosseini

    Abstract: Transformer models are increasingly used for whole-slide image analysis in computational pathology. Yet, WSIs differ fundamentally from natural images: neighbouring patches often contain highly similar tissue type, stain, texture, and cellular composition. We identify this local spatial redundancy as a pathology-specific failure mode of self-attention, where dominant neighbourhood features can be… ▽ More

    Submitted 12 August, 2026; v1 submitted 8 August, 2026; originally announced August 2026.

    Comments: Accepted at BMVC 2026 Conference

  16. arXiv:2608.08290  [pdf, ps, other

    cs.CV

    Test-Time Prototype Adaptation for Open-Vocabulary Semantic Segmentation

    Authors: Haozhe Wang, Jintao Cheng, Weibin Li, Xiaoyu Tang

    Abstract: Open-vocabulary semantic segmentation (OVSS) repurposes a pretrained CLIP encoder for dense prediction without additional labeled supervision. Existing methods improve CLIP's spatial behavior either by redesigning its internal attention or by injecting features from auxiliary vision foundation models; both require access to the host's internal computation and are tailored to its specific forward p… ▽ More

    Submitted 8 August, 2026; originally announced August 2026.

    Comments: 17 pages, 12 figures, preprint

    ACM Class: I.2.10; I.4.6

  17. arXiv:2608.08097  [pdf, ps, other

    cs.DC

    OasisKV: Scaling In-Decode KV Cache Beyond HBM with Lookahead Sparse Prefetching

    Authors: Can Xiao, Sukmin Cho, Junbong We, Zhixiong Niu, Jianyi Cheng, Yiren Zhao, Youngjin Kwon, Yongqiang Xiong, Rui Ma, Junyi Liu

    Abstract: Large language model (LLM) inference serving is increasingly constrained by memory rather than compute. As long-context and long-form reasoning workloads become more prevalent, the key-value (KV) cache dominates both memory footprint and memory traffic during LLM token generation, i.e., decode. In particular, HBM capacity has become a scarce and costly resource that heavily limits inference batch… ▽ More

    Submitted 8 August, 2026; originally announced August 2026.

  18. arXiv:2608.06878  [pdf, ps, other

    cs.CV

    ControlRef: Efficient Layout-Guided Multi-Instance Generation via Anchored 4D-RoPE

    Authors: Yunkai Yang, Yudong Zhang, Xinying Chen, Haoyuan Liang, Yizhuo Niu, Jinshuai Cheng, Kunquan Zhang, Liziyue Fang, Weitao Wan, Runmin Dong

    Abstract: Layout-guided multi-instance generation is essential for controllable image synthesis in Multi-Modal Diffusion Transformers (MM-DiTs). However, integrating this capability into unified architectures remains challenging. Prior frameworks rely on redundant full-resolution canvas padding and Shifted-RoPE to manage multiple reference images. This mechanism drastically inflates computational overhead f… ▽ More

    Submitted 7 August, 2026; originally announced August 2026.

  19. arXiv:2608.06794  [pdf, ps, other

    cs.CV

    PAST: Prompt-Adaptive Sampling Termination for Efficient Diffusion Model

    Authors: Renye Yan, Jikang Cheng, You Wu, Wei Peng, Zongwei Wang, Ling Liang, Yimao Cai

    Abstract: While diffusion models have made significant progress in text-to-image tasks, they still exhibit limitations when directly optimizing downstream objectives. Although Reinforcement Learning (RL) enables targeted optimization, existing methods are generally constrained by low-efficiency fine-tuning and sparse rewards. To address these challenges, we propose PAST, which provides differentiated reward… ▽ More

    Submitted 7 August, 2026; originally announced August 2026.

  20. arXiv:2608.06768  [pdf, ps, other

    cs.CV

    Explore or Converge? Stage-Guided Per-Step Optimization for Diffusion Models

    Authors: Renye Yan, Jikang Cheng, You Wu, Wei Peng, Zongwei Wang, Ling Liang, Yimao Cai

    Abstract: Diffusion models have strong generative capabilities. However, their maximum likelihood training objective only focuses on reconstructing the data distribution, making it difficult to align with specific preferences. Reinforcement learning (RL) for preference alignment in diffusion models is promising but limited by reward sparsity. Since a single reward cannot support optimization, existing RL me… ▽ More

    Submitted 6 August, 2026; originally announced August 2026.

  21. arXiv:2608.06697  [pdf

    physics.chem-ph

    3D Molecular Representation Learning for Organic Mixtures: Viscosity and Density Prediction

    Authors: Haicheng Qu, Yanyi Su, Ning Wang, Shangqian Chen, Zhifeng Gao, Jun Cheng, Qi Ou

    Abstract: The viscosity and density of organic mixtures are essential properties for designing lubricants, solvents, and heat transfer fluids. In engineering practice, formulating a functional fluid requires understanding how these properties change with composition and temperature. However, exhaustive experimental characterization across the full parameter space is impractical due to the vast number of pos… ▽ More

    Submitted 6 August, 2026; originally announced August 2026.

  22. arXiv:2608.06109  [pdf, ps, other

    hep-ex

    Search for the charged lepton flavour violating decay $η'\to eμ$

    Authors: BESIII Collaboration, M. Ablikim, M. N. Achasov, P. Adlarson, X. C. Ai, C. S. Akondi, R. Aliberti, A. Amoroso, Q. An, Y. H. An, M. S. Anderson, Y. Bai, O. Bakina, H. R. Bao, X. L. Bao, M. Barbagiovanni, V. Batozskaya, K. Begzsuren, N. Berger, M. Berlowski, M. B. Bertani, D. Bettoni, F. Bianchi, E. Bianco, A. Bortone , et al. (744 additional authors not shown)

    Abstract: Based on $(8998\pm40)\times10^6$ $J/ψ$ events collected in $e^+e^-$ collisions at $\sqrt{s} = 3.097$ GeV with the BESIII detector, we present a search for the charged lepton flavour violating decay $η'\to eμ$ with $J/ψ\toγη'$. No significant signal is observed, and an upper limit on its decay branching fraction is set to be $6.3\times10^{-7}$ at the 90% confidence level, improving the previous bes… ▽ More

    Submitted 6 August, 2026; originally announced August 2026.

    Comments: 20 pages, 5 figures, 3 tables

  23. arXiv:2608.04893  [pdf, ps, other

    cs.CR cs.AI cs.LG

    When Does Latent Communication Pay? A Causal Audit of Relayed KV Caches in Multi-Agent LLMs

    Authors: Jiaming Cheng, Subhransu Das, Rajiv Ramnath

    Abstract: Multi-agent LLM systems relay key--value caches instead of text and credit their gains to exchanged ``latent thoughts''. That credit is a claim about \emph{which} example's cache is relayed, not merely that one is. We audit it causally in released systems. The cache is replaced with deranged (mismatched-example), zeroed, and moment-matched random counterparts, under two regimes defined by whethe… ▽ More

    Submitted 5 August, 2026; originally announced August 2026.

  24. arXiv:2608.04633  [pdf, ps, other

    cs.RO

    Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models

    Authors: Xingyu Ding, Yuzhong Zhao, Yang Wu, Chaoyang Zhao, Chunhai Zhao, Yifan Zhang, Jian Cheng

    Abstract: Recent Vision-Language-Action (VLA) methods improve generalization by aligning their representations with 3D scene geometry. However, these methods are fundamentally instruction-agnostic: the representations align the entire scene uniformly, neglecting the 3D geometry of the specific target object designated by the language instruction. This causes failures on fine-grained manipulation and target… ▽ More

    Submitted 5 August, 2026; originally announced August 2026.

  25. arXiv:2608.04333  [pdf, ps, other

    cs.LG

    Cost-Aware Multi-Objective Bandits: Theory and Application to Budgeted LLM Configuration Evaluation

    Authors: Bo Xue, Zhi Hong, Jiayi Li, Yuanyu Wan, Ji Cheng, Shuang Qiu

    Abstract: Large language model (LLM) configuration evaluation is challenging due to limited evaluation budgets, varying costs, and multiple competing objectives. In this paper, we formulate LLM configuration evaluation as a cost-aware multi-objective bandit problem, where each configuration evaluation incurs a configuration-dependent cost and yields a noisy vector-valued outcome. Under this framework, we st… ▽ More

    Submitted 4 August, 2026; originally announced August 2026.

  26. arXiv:2608.04324  [pdf, ps, other

    cs.LG cs.AI

    Efficient Online Lexicographic Generalized Low-Rank Matrix Bandits

    Authors: Bo Xue, Ji Cheng, Haodong Jing, Hongzong Li, Shuang Qiu

    Abstract: This paper studies generalized low-rank matrix bandits with multiple prioritized objectives. At each round, the learner selects a matrix-valued arm and observes a vector-valued reward, whose components correspond to multiple objectives with different priority levels. Each objective is governed by an objective-specific generalized low-rank matrix model, and the learner evaluates arms according to a… ▽ More

    Submitted 4 August, 2026; originally announced August 2026.

  27. arXiv:2608.04128  [pdf, ps, other

    quant-ph physics.atom-ph

    Quantum Contextuality and Entanglement-Free Grover Search in a Trapped-Ion Optical Qudit

    Authors: Tarun Dutta, Jasper Phua Sing Cheng, Alex Jin, Sergi Ramos-Calderer, José Ignacio Latorre, Manas Mukherjee

    Abstract: Quantum computational advantage is generally attributed to coherent interference and other non-classical resources, yet their respective roles remain difficult to disentangle in experimental platforms where multipartite entanglement is inherently present. High-dimensional quantum systems provide an attractive route for investigating these resources while simultaneously reducing hardware overhead f… ▽ More

    Submitted 4 August, 2026; originally announced August 2026.

    Comments: 12 pages, 10 figures

  28. arXiv:2608.03930  [pdf, ps, other

    cs.CL cs.AI cs.LG

    Logic Before Language: Pre-pretraining on Formal Derivations Fosters Skill Acquisition and Compressibility

    Authors: Jo-Ku Cheng, Nikolaos Aletras, Marco Valentino

    Abstract: Pre-pretraining language models (LMs) on symbolic data can accelerate and improve natural language acquisition. However, existing pre-pretraining tasks, such as Dyck and procedural algorithms, rely on narrow primitives that fail to capture the expressive capacity of natural language. Moreover, prior studies remain restricted to relatively small token budgets, offering limited insight into skill em… ▽ More

    Submitted 4 August, 2026; originally announced August 2026.

  29. arXiv:2608.03138  [pdf, ps, other

    cs.CL cs.AI

    Internalizing Academic Writing Workflows for Introduction Generation via Struct-Aware Policy Learning

    Authors: Meicong Zhang, Tiancheng Su, Jiahao Cheng, Guoxiu He, Xinqi Tao, Dejia Song

    Abstract: Generating a rigorous paper introduction with large language models (LLMs) remains challenging, since it requires coordinating background, gap identification, method and contribution within a coherent narrative. Existing solutions externalize this process as multi-stage prompts or agent workflows which are expensive and vulnerable to cross-stage drift. We propose StructPO, a struct-aware policy le… ▽ More

    Submitted 4 August, 2026; originally announced August 2026.

  30. arXiv:2608.02099  [pdf, ps, other

    cs.AR cs.AI cs.CV

    DeGS: A Scalable 3DGS Architecture via Decoupled Workload Parsing and Reorganization

    Authors: Minnan Pei, Gang Li, Zeyu Zhu, Siting Wang, Junwen Si, Zhuoran Song, Yu Feng, Fangxin Liu, Xiaoyao Liang, Jian Cheng

    Abstract: 3D Gaussian Splatting (3DGS) has emerged as a leading technique for real-time novel view synthesis, yet existing 3DGS accelerators suffer from poor architectural scalability: increasing the number of PEs leads to marginal performance improvement during rendering. We identify that the root cause is the tightly coupled ``checking-while-blending'' dataflow, which exacerbates PE underutilization cause… ▽ More

    Submitted 3 August, 2026; originally announced August 2026.

    Comments: Accepted to the 59th IEEE/ACM International Symposium on Microarchitecture (MICRO 2026)

  31. arXiv:2608.01973  [pdf, ps, other

    cs.RO cs.CV

    Roomer: Reflective Object-Grounded Model Editing and Repair for 3D Indoor Layout Synthesis

    Authors: Lingwei Dang, Ziyan Qiu, Jiajia Cheng, Shishuo Shang, Zhenhao Zhang, Yufei Zhu, Qingxin Xiao, Pan Liu, Shenghui Huang, Yun Hao, Juntong Li, Qingyao Wu

    Abstract: Existing indoor layout generators produce globally plausible layouts yet may retain local violations such as collisions, out-of-bounds placements, obstructed openings, and blocked circulation. Most prior work focuses on full-scene synthesis or scene-level optimization, with limited support for identifying responsible objects and locally repairing affected regions. We present Roomer, a reflective r… ▽ More

    Submitted 3 August, 2026; originally announced August 2026.

  32. arXiv:2608.01954  [pdf, ps, other

    cs.CV

    StyleForge: Indoor Furniture Styling by Counterfactual Reasoning in a Hypergraph Field

    Authors: Lingwei Dang, Shishuo Shang, Pan Liu, Jiajia Cheng, Ziyan Qiu, Zhenhao Zhang, Yufei Zhu, Shenghui Huang, Qingxin Xiao, Yun Hao, Juntong Li, Qingyao Wu

    Abstract: Fixed-layout indoor furniture styling requires selecting assets that form a coherent room without changing the prescribed furniture categories, positions, orientations, or scales. Existing approaches typically retrieve each asset independently or rely on static local relations, making them prone to shape, material, and color conflicts after scene composition. We introduce StyleForge, a scene-level… ▽ More

    Submitted 3 August, 2026; originally announced August 2026.

  33. arXiv:2608.01410  [pdf, ps, other

    cs.RO cs.CV

    GenTrack: Physical Alignment for Robot-Native Motion Generation and Zero-Shot Humanoid Tracking

    Authors: Zeyu Ling, Xinyao Yu, Renye Yan, Jikang Cheng, Zhanke Wang, Qing Shuai, Changqing Zou

    Abstract: General-purpose humanoid trackers can execute diverse references, but their zero-shot coverage depends on large embodied corpora that are costly to extend. Text-to-motion generators offer scalable supervision, yet models trained on human motion or retargeted data inherit a gap between kinematic plausibility and robot executability. Existing one-way pipelines fix either the generated corpus or the… ▽ More

    Submitted 5 August, 2026; v1 submitted 2 August, 2026; originally announced August 2026.

  34. arXiv:2608.01269   

    cs.CL cs.AI

    ACE-GraphRAG: Agentic Context Engineering for Hierarchical GraphRAG

    Authors: Yongfeng Huang, Yuren Lai, Ruiying Chen, Haoyu Huang, Mingming Zhao, James Cheng

    Abstract: Hierarchical Graph Retrieval-Augmented Generation (GraphRAG) organizes corpus knowledge at multiple levels of granularity, yet fixed context construction may fail to translate these multi-resolution representations into a context suited to the current query. We identify this mismatch as the representation--inference gap. We propose Agentic Context Engineering for Hierarchical GraphRAG (ACE-GraphRA… ▽ More

    Submitted 3 August, 2026; v1 submitted 2 August, 2026; originally announced August 2026.

    Comments: Withdrawn because the manuscript was posted prematurely before completion of the required internal review and release authorization

  35. arXiv:2608.01266  [pdf

    cond-mat.mtrl-sci

    Why Ammoniated Lithium Borohydrides Liquefy and Resolidify?

    Authors: Qian Wang, Zixin Xu, Ryuhei Sato, Hiroki Miyaoka, Takayuki Ichikawa, Eric Jianfeng Cheng, Shin-ichi Orimo, Fangqin Guo, Hao Li

    Abstract: Ammonia ($\mathrm{NH_3}$) absorption drives $\mathrm{LiBH_4\!\cdot\!xNH_3}$ through a re-entrant ``solid--liquid--solid'' transition: $\mathrm{LiBH_4\!\cdot\!NH_3}$ is a well-defined solid ammoniate, compositions near $\mathrm{LiBH_4\!\cdot\!2NH_3}$ are liquid-like or partially liquefied, whereas $\mathrm{LiBH_4\!\cdot\!3NH_3}$ returns to a more rigid non-liquid ammoniate state. However, the micro… ▽ More

    Submitted 3 August, 2026; v1 submitted 2 August, 2026; originally announced August 2026.

    Comments: 21 pages, 4 figures

  36. arXiv:2608.00779  [pdf, ps, other

    cs.RO

    SIPTraj: Map-Free End-to-End Trajectory Prediction via Physics-Guided Scene Interaction

    Authors: Feifei Liu, Zejun Wei, Haozhe Wang, Yazhi Ye, Yuying Zhang, Jintao Cheng, Chi Man Vong, Xieyuanli Chen, Xiaoyu Tang

    Abstract: Trajectory prediction of surrounding agents is a prerequisite for safe planning and decision making in autonomous driving. Without high-definition (HD) maps, sensor-derived bird's-eye-view (BEV) features provide no explicit lane topology or drivable-area priors, making it inherently difficult to ground each agent in its surrounding scene context. Moreover, physical feasibility remains difficult to… ▽ More

    Submitted 1 August, 2026; originally announced August 2026.

  37. arXiv:2607.29600  [pdf, ps, other

    cs.RO

    HAM-VLN: Harnessing Hierarchical Agentic Memory for Zero-Shot Vision-and-Language Navigation

    Authors: An Liu, Bingxi Liu, Hongyu Ding, Yixuan Jiang, Yaran Chen, Fulin Tang, Cong Leng, Hong Zhang, Jian Cheng

    Abstract: Vision-and-language navigation (VLN) enables robots to follow instructions in previously unseen environments. Recently, a training-free paradigm has emerged: the robot queries a multimodal LLM to understand its observations and plan the next action. However, long-horizon navigation based on either image streams or dense map inevitably introduces a growing memory and reasoning bottleneck. We presen… ▽ More

    Submitted 31 July, 2026; originally announced July 2026.

  38. arXiv:2607.29561  [pdf, ps, other

    cs.LG cs.AI

    MOT-SR: Multi-Objective Tool-Augmented Scientific Equation Discovery with Large Language Models

    Authors: Boxiao Wang, Runxiang Wang, Kai Li, Chongming Li, Zhiwei Chen, Yifan Zhang, Jian Cheng

    Abstract: Symbolic Regression (SR) aims to discover analytical equations from observational data and plays a central role in scientific modeling. While recent Large Language Model (LLM) based approaches show promise, they face two limitations. First, they lack data analysis mechanisms for uncovering variable dependencies, which reduces the efficiency of equation discovery. Second, most methods rely on singl… ▽ More

    Submitted 31 July, 2026; originally announced July 2026.

    Comments: Code is available at https://github.com/wswbx/MOT-SR

  39. arXiv:2607.29222  [pdf, ps, other

    cs.CV

    Is It Time for the Renaissance of Salient Object Detection in the Era of MLLMs?

    Authors: Wenzhuo Zhao, Xiuzhi Li, Zhongkuan Mao, Ronghao Xian, Yao Jiang, Zhao Gao, Keren Fu, Qijun Zhao, Jian Cheng

    Abstract: The zero-shot capabilities of multimodal large language models (MLLMs) are pushing salient object detection (SOD) beyond task-specific supervision. To disentangle MLLMs beyond conventional mask-based evaluation, we decompose SOD into localization and segmentation, and re-engineer datasets with phrases, boxes, and attributes, establishing a diagnostic benchmark for MLLM saliency perception (SaliLLM… ▽ More

    Submitted 31 July, 2026; originally announced July 2026.

    Comments: 10 pages, 4 figures, conference

  40. arXiv:2607.27953  [pdf, ps, other

    cs.LG

    AutoPref: Automatic Discovery of Task-Specific Preference Objectives for Neural Combinatorial Optimization

    Authors: Shengda Gu, Kai Li, Xinyi Ke, Haobo Fu, Yifan Zhang, Jian Cheng

    Abstract: Combinatorial optimization problems (COPs) underpin many real-world decisions, but their exponentially large search spaces make high-quality solutions costly to obtain. Neural combinatorial optimization (NCO) learns fast construction policies, typically with reinforcement learning (RL), while preference-based NCO improves sample efficiency by learning from relative solution quality. However, exist… ▽ More

    Submitted 30 July, 2026; originally announced July 2026.

    Comments: 8pages, 2figures

  41. arXiv:2607.26712  [pdf, ps, other

    cs.RO

    ActSWM: Action-Sensitive World Models for Long-Horizon Planning in Open-World Games

    Authors: Zhenfeng Gan, ZiTong Zeng, Jiajun Cheng, Yeke Song, Yongyi Tang, Xueqian Wang

    Abstract: Latent world models support efficient model-predictive control by optimizing future control sequences in latent space and replanning in a receding-horizon manner. However, existing latent predictors often lack stable long-horizon rollout ability, and prediction accuracy alone does not ensure that rollouts remain responsive to the actions being planned. We identify Context Collapse, a failure mode… ▽ More

    Submitted 15 August, 2026; v1 submitted 29 July, 2026; originally announced July 2026.

    Comments: 10 pages, 5 figures

  42. arXiv:2607.26490  [pdf, ps, other

    cs.AI cs.LG cs.NE

    EvoPINN: Agentic Discovery of Executable Algorithms for Physics-Informed Neural Networks

    Authors: Peng Yin, Kai Li, Yifan Zhang, Jian Cheng

    Abstract: Physics-informed neural networks (PINNs) have emerged as a powerful paradigm for solving partial differential equations (PDEs), yet their performance heavily relies on the manual, trial-and-error engineering of neural representations, loss formulations, and optimization dynamics. While Large Language Models (LLMs) offer a promising avenue for automated design, unconstrained code generation often y… ▽ More

    Submitted 29 July, 2026; originally announced July 2026.

  43. Instruction-based Image Editing: A Survey on Data, Models, Evaluation, and Applications

    Authors: Xianghao Zang, Zijian Jiang, Jiarong Cheng, Qianrui Teng, Ying He, Yuxuan Mu, Chao Ban, Huayu Zhang, Lanxiang Zhou, Zerun Feng, Chi Zhang

    Abstract: Instruction-based Image Editing (IIE) aims to transform a given image into a new one based on textual instructions. Advances in Large Language Models (LLMs) and Vision-Language Models (VLMs) have accelerated progress toward practical ``one-sentence image editing" systems. This survey presents a systematic taxonomy and comprehensive review of IIE research, structured around five core dimensions: (1… ▽ More

    Submitted 28 July, 2026; originally announced July 2026.

    Comments: 33 pages, 7 figures, Vicinagearth

    Journal ref: Zang, X., Jiang, Z., Cheng, J. et al. Instruction-based image editing: a survey on data, models, evaluation, and applications. Vicinagearth 3, 3 (2026)

  44. arXiv:2607.25596  [pdf, ps, other

    nlin.SI

    Tau functions of the constrained matrix KP hierarchy

    Authors: Xiaohan Fan, Jipeng Cheng, Jinbiao Wang

    Abstract: The constrained matrix KP hierarchy $(L^{k})_{<0}=\sum_{i=1}^{m}Q_{i}\partial^{-1}R_{i}^{\intercal}$ is investigated from the aspects of tau functions. Firstly, the matrix KP hierarchy is viewed as one special reduction of the multi-component KP hierarchy. Then bilinear equations of the constrained matrix KP hierarchy as the multi-component KP hierarchy are given in terms of tau functions. Finally… ▽ More

    Submitted 28 July, 2026; originally announced July 2026.

    Comments: 22 pages

    MSC Class: 35C08; 35Q53; 37K10; 37K40

  45. arXiv:2607.25157  [pdf, ps, other

    cs.AI

    PreDiff-LM: Pretrained Discrete Masked Diffusion Language Modeling with Hybrid Attention

    Authors: Zhengtao Yao, Runhao Li, Xupeng Chen, Jiayi Cheng, Chenqian Le, Michael Yue, Jesson Wang, Siheng Wang, Guang Yang, Haoyan Xu, Chenhao Wei, Zhengqing Yuan, Youran Shen, Yanfang Ye, Junhao Dong

    Abstract: Discrete masked diffusion language models support bidirectional generation and infilling, but adapting pretrained autoregressive (AR) transformers requires reconciling causal pretraining with bidirectional denoising. We study this problem at the level of attention rather than claiming AR-weight reuse itself as novel. PreDiff-LM preserves causal attention within the observed prompt while allowing f… ▽ More

    Submitted 27 July, 2026; originally announced July 2026.

  46. arXiv:2607.25136  [pdf, ps, other

    cs.AI

    Less Data, Better Alignment: Data-Centric Multi-Evaluator Agreement for Preference Optimization

    Authors: Zhengtao Yao, Runhao Li, Xupeng Chen, Jiayi Cheng, Chenqian Le, Michael Yue, Siheng Wang, Haoyan Xu, Yuqi Li, Chenhao Wei, Zhengdao Li, Rongchao Zhang, Guang Yang, Yidong Wang, Junhao Dong

    Abstract: Research on preference optimization often varies the training objective while holding the data fixed. We instead ask whether a small, high-confidence set of on-policy responses can provide a reliable learning signal. Our method, DMAPO (Data-centric Multi-evaluator Agreement for Preference Optimization), generates candidate responses from the target policy, evaluates helpfulness, factuality, and co… ▽ More

    Submitted 27 July, 2026; originally announced July 2026.

    Comments: 19 pages

  47. arXiv:2607.24869  [pdf, ps, other

    cs.IR cs.CL cs.LG

    Ranked by Position: Order Sensitivity as an Exploitable Attack Surface in LLM Listwise Recommenders

    Authors: Ge Zhang, Jingru Cheng, Huiyuan Chen

    Abstract: Large language models (LLMs) used as listwise rerankers in recommendation systems suffer from position bias when serializing candidate sets into prompts. We show this order sensitivity creates an exploitable attack surface: an attacker can promote a label-0 target into the top-$k$ solely by reordering candidates, without changing item content, labels, or model parameters. We introduce… ▽ More

    Submitted 26 July, 2026; originally announced July 2026.

    Comments: 13 pages, 6 figures, 11 tables. Code and data available at https://github.com/geoz-lab/position_bias_attack

  48. arXiv:2607.24862  [pdf, ps, other

    cs.IR

    KuaiLive-M3: A Multi-Modal, Multi-Domain, and Multi-Feedback Dataset for Live Streaming Recommendation

    Authors: Ke Guo, Changle Qu, Jiayaqi Cheng, Xiao Zhang, Shijun Wang, Xiaoyu Zhang, Xueliang Wang, Le Zhang, Lantao Hu, Jun Xu

    Abstract: Existing public live streaming datasets suffer from three major limitations: they provide limited access to temporally evolving multimodal live content, overlook users' cross-domain interactions between short videos and live streams, and contain only implicit behavioral signals without explicit feedback that captures users' perceived content quality and satisfaction. These limitations prevent exis… ▽ More

    Submitted 26 July, 2026; originally announced July 2026.

  49. arXiv:2607.24531  [pdf, ps, other

    astro-ph.HE

    R-process nucleosynthesis from magnetar giant flares in neutron star--white dwarf mergers: A unified picture for peculiar long gamma-ray bursts

    Authors: Shu-Qing Zhong, Long Li, Le Zou, Ji-Gui Cheng, Yan-Zhi Meng, Jia-Hong Gu

    Abstract: Peculiar long gamma-ray bursts (GRBs), exemplified by GRBs 211211A and 230307A, exhibit a long-duration multi-component prompt emission, an X-ray plateau in their afterglow, and a kilonova signature. Their origin remains highly debated. In this work, we present a unified picture for these events based on neutron star--white dwarf (NS--WD) mergers involving a pre-merger magnetar and a massive WD. I… ▽ More

    Submitted 27 July, 2026; originally announced July 2026.

    Comments: accepted by A&A

  50. arXiv:2607.23970   

    cs.LG cs.AI cs.CL

    Understanding Machine Unlearning Through the Lens of Mode Connectivity

    Authors: Jiali Cheng, Hadi Amiri

    Abstract: Machine Unlearning aims to remove undesired information from trained models without full retraining from scratch. Despite recent progress, the loss landscape and optimization geometry of unlearning are poorly understood. In this paper, we study machine unlearning through the lens of mode connectivity--the phenomenon that independently trained models can often be connected by smooth low-loss paths… ▽ More

    Submitted 3 August, 2026; v1 submitted 26 July, 2026; originally announced July 2026.

    Comments: This work was intended as a replacement of arXiv:2504.06407 and any subsequent updates will appear there