Skip to content

March 2026 arXiv papers — page 105

Showing 10,40110,500 of 25,974 papers

  1. Hien Duy Nguyen

    We prove that finite multivariate Erlang mixture densities with a common rate parameter are dense in the class of probability densities on $\mathbb{R}_{+}^{d}$ that belong to $L^{p}$, for every dimension $d\in\mathbb{N}$ and every $1\le p<\infty$. The argument is constructive: the one-dimensional Sz\'asz--Mirakjan--Kantorovich operator yields Erlang mixture

  2. Jingzhi Chen, Lijian Xu

    The protein folding problem has been fundamentally transformed by artificial intelligence, evolving from static structure prediction toward the modeling of dynamic conformational ensembles and complex biomolecular interactions. This review systematically examines the paradigm shift in AI driven protein science across five interconnected dimensions: unified m

  3. Julian Martinez, Kooktae Lee

    Efficient coordination for collective spatial distribution is a fundamental challenge in multi-agent systems. Prior research on Density-Driven Optimal Control (D2OC) established a framework to match agent trajectories to a desired spatial distribution. However, implementing this as a predictive controller requires solving a large-scale Karush-Kuhn-Tucker (KK

  4. Teerapong Panboonyuen

    Automated property risk detection is a high-impact yet underexplored frontier in computer vision with direct implications for real estate, underwriting, and insurance operations. We introduce HOMEY (Heuristic Object Masking with Enhanced YOLO), a novel detection framework that combines YOLO with a domain-specific masking mechanism and a custom-designed loss

  5. Mingde Zhou, Zheng Chen, Yulun Zhang

    Video compression aims to maximize reconstruction quality with minimal bitrates. Beyond standard distortion metrics, perceptual quality and temporal consistency are also critical. However, at ultra-low bitrates, traditional end-to-end compression models tend to produce blurry images of poor perceptual quality. Besides, existing generative compression methods

  6. Kazuaki Hashiyama, Takeshi Nakamori, Anju Sato, Mana Hasebe

    We report our optical observations of the Crab pulsar using the Imager of MPPC-based Optical photoN counter from Yamagata (IMONY), a high-time-resolution photon-counting imager with 100 ns timing resolution, mounted on the 3.8 m Seimei telescope in Japan (f/D~6). The detector format was upgraded from a $4\times4$ to an $8\times8$ GAPD array with larger pixel

  7. Yuanze Ding, Michael J. Koss, Fiona A. Harrison, Charles C. Steidel

    We present high spatial resolution ($\lesssim$1.0''), multi-wavelength observations of UGC 2369S, a nearby luminous infrared galaxy showing three distinct cores separated on kpc scales in near-infrared (NIR) imaging with significant X-ray emission. Utilizing optical/NIR adaptive optics (AO), radio, \chandra X-ray, as well as archival HST imaging, we perform

  8. Mengting Shen, Jun Yin, Hassen M. Yesuf, Lei Hao

    We analyze a clean sample of 1,118 late-type, face-on galaxies without AGN contamination from the MaNGA survey. Their photometric structures are quantified via two-component (bulge+disk) decompositions on deep $g$-band images from the DESI Legacy Survey. Using a disk central surface brightness of $\mu_{\rm 0,d,cor}$(g) = 22 $\pm$ 0.3 mag arcsec$^{-2}$ (corre

  9. Quilee Simeon

    Inferring the connectivity of neural circuits from incomplete observations is a fundamental challenge in neuroscience. We present a covariance-based method for estimating the weight matrix of a recurrent neural network from sparse, partial measurements across multiple recording sessions. By accumulating pairwise covariance estimates across sessions where dif

  10. Daniel DeTone, Federica Bogo, Eric-Tuan Le, Duncan Frost

    The Nymeria Dataset, released in 2024, is a large-scale collection of in-the-wild human activities captured with multiple egocentric wearable devices that are spatially localized and temporally synchronized. It provides body-motion ground truth recorded with a motion-capture suit, device trajectories, semi-dense 3D point clouds, and in-context narrations. In

  11. Jooyoung Kim, Wonje Choi, Younguk Song, Honguk Woo

    Recent advances in Vision-Language Models (VLMs) have enabled video-instructed robotic programming, allowing agents to interpret video demonstrations and generate executable control code. We formulate video-instructed robotic programming as a cross-domain adaptation problem, where perceptual and physical differences between demonstration and deployment induc

  12. Yao Zhang, Yuchen Song, Xiao Luo, Shengnan Li

    Recent advances in large language models (LLMs) have demonstrated strong capabilities in code generation and text synthesis, yet their potential for symbolic physical reasoning in domain-specific scientific problems remains underexplored. We present a mathematical reasoning enhanced generative AI approach for optical communication formula derivation, focusin

  13. Seonghyun Jin, Jong Chul Ye

    Streaming 3D reconstruction maintains a persistent latent state that is updated online from incoming frames, enabling constant-memory inference. A key failure mode is the state update rule: aggressive overwrites forget useful history, while conservative updates fail to track new evidence, and both behaviors become unstable beyond the training horizon. To add

  14. Deepak Kumar, S Pushpavanam

    Tear film rupture on the corneal surface plays a critical role in ocular health and visual comfort. Conventional theoretical approaches often idealize the cornea as a perfectly smooth surface, ignoring the surface roughness that are characteristic of healthy as well as diseased eyes. In this study, we develop a comprehensive mathematical model to investigate

  15. Yiqi Luo, Xue Luo

    We investigate Bayesian nonparametric density estimation via orthogonal polynomial expansions in weighted Sobolev spaces. A core challenge is establishing minimax optimal posterior convergence rates, especially for densities on unbounded domains without a strictly positive lower bound. For densities bounded away from zero, we give sufficient conditions under

  16. Minsoo Cheong, Donghyun Son, Woosang Lim, Sungjoo Yoo

    Diffusion-based large language models (dLLMs) rely on bidirectional attention, which prevents lossless KV caching and requires a full forward pass at every denoising step. Existing approximate KV caching methods reduce this cost by selectively updating cached states, but their decision overhead scales with context length or model depth. We propose EntropyCac

  17. Bo Zhao, Yihang Liu, Chenfeng Zhang, Huan Yang

    Text-guided texture editing aims to modify object appearance while preserving the underlying geometric structure. However, our empirical analysis reveals that even SOTA editing models frequently struggle to maintain structural consistency during texture editing, despite the intended changes being purely appearance-related. Motivated by this observation, we j

  18. Hao-Hao Chen, Wen-Tao Xu, Xin-Yu Liang, Ming-Xuan Lu

    The physical origin of fast radio bursts (FRBs) remains an unsolved mystery in astrophysics, with the magnetar central engine model as the leading framework. Systematically searching for physical associations between FRBs and the energetic astrophysical transients (ATs) that form magnetars provides a critical test of this scenario, and key clues to FRB proge

  19. Jing Liu, Zhenchao Ma, Han Yu, Bobo Ju

    Recent advances in collaborative knowledge distillation have demonstrated cutting-edge performance for resource-constrained distributed multimedia learning scenarios. However, achieving such competitiveness requires addressing a fundamental mismatch: high-dimensional teacher knowledge complexity versus heterogeneous client learning capacities, which currentl

  20. Jing Liu, Zhengliang Guo, Yan Wang, Xiaoguang Zhu

    Federated learning (FL) is severely challenged by non-independent and identically distributed (non-IID) client data, a problem that degrades global model performance, especially in multimodal perception settings. Conventional methods often fail to address the underlying semantic discrepancies between clients, leading to suboptimal performance for multimedia

  21. Takumi Oshima, Yamato Arai, Koji Hukushima

    Phase transitions in a modified Nishimori model, including the model considered by Kitatani, on a two-dimensional square lattice are investigated using a tensor-network-based sampling scheme. In this model, generating bond configurations is computationally demanding because of the correlated random interactions. The employed sampling method enables hierarchi

  22. Siqi Song, Fulin Wu, Zhong-Qiu Wang

    Due to the absence of clean reference signals and spatial cues, monaural unsupervised speech dereverberation is a challenging ill-posed inverse problem. To realize it, we propose augmented reverberant-target training (ARTT), which consists of two stages. In the first stage, reverberant-target training (RTT) is proposed to first further reverberate the observ

  23. Omar Astudillo-Marbán, Oriol Solé-Pi

    Given a set P of points on the plane, a polygon with vertices in P is said to be empty if it contains no element of P in its interior. We show that every set of n points in general position on the plane determines at least $\Omega(n^{20/11})$ empty convex pentagons (also known as 5-holes). This result improves upon the previous bound of $\Omega(n\cdot(\log n

  24. Reza Ghane, Danil Akhtiamov, Babak Hassibi

    In the present paper we study the performance of linear denoisers for noisy data of the form $\mathbf{x} + \mathbf{z}$, where $\mathbf{x} \in \mathbb{R}^d$ is the desired data with zero mean and unknown covariance $\mathbf{\Sigma}$, and $\mathbf{z} \sim \mathcal{N}(0, \mathbf{\Sigma}_{\mathbf{z}})$ is additive noise. Since the covariance $\mathbf{\Sigma}$ is

  25. Ziyi Wang, Qizan Guo, Rishitosh Singh, Xiyang Hu

    Inferring human engagement from gameplay video is important for game design and player-experience research, yet it remains unclear whether vision--language models (VLMs) can infer such latent psychological states from visual cues alone. Using the GameVibe Few-Shot dataset across nine first-person shooter games, we evaluate three VLMs under six prompting stra

  26. Ryota Kojima

    The criticality hypothesis posits that biological neural networks operate near a phase transition, yet within standard Gaussian mean-field theories this regime appears fragile and requires fine tuning. Here we show that heavy-tailed synaptic connectivity provides a robust alternative mechanism. By developing a dynamical mean-field theory for Cauchy-distribut

  27. Chunhao Liao, Hongxu Xu, Xintong Zhou, Zhenyang Xu

    Peephole optimizations are a core component of modern optimizing compilers. It rewrites specific instruction into semantically equivalent but more efficient forms. In practice, creating a new peephole optimization often starts from a concrete optimization instance and requires lifting it into a more general rewrite rule that matches a wider range of instruct

  28. Haosen Li, Qi Meng, Jiahao Li, Rui Zhang

    Partial differential equation (PDE) simulation holds extensive significance in scientific research. Currently, the integration of deep neural networks to learn solution operators of PDEs has introduced great potential. In this paper, we present UniFluids, a conditional flow-matching framework that harnesses the scalability of diffusion Transformer to unify l

  29. Xu'an Dou, Louis Tao, Zhe Xue, Zhennan Zhou

    In large-scale excitatory neuronal networks, rapid synchronization manifests as {multiple firing events (MFEs)}, mathematically characterized by a finite-time blow-up of the neuronal firing rate in the mean-field Fokker-Planck equation. Standard numerical methods struggle to resolve this singularity due to the divergent boundary flux and the instantaneous na

  30. Haonan Yu, Junhao Liu, Zhenyu Yan, Haoran Lin

    Precise behavioral control of large language models (LLMs) is critical for complex applications. However, existing methods often incur high training costs, lack natural language controllability, or compromise semantic coherence. To bridge this gap, we propose WASD (unWeaving Actionable Sufficient Directives), a novel framework that explains model behavior by

  31. Zehaan Naik, Debasis Kundu

    Least Absolute Deviations (LAD) regression provides a robust alternative to ordinary least squares by minimizing the sum of absolute residuals. However, its widespread use has been limited by the computational cost of existing solvers, particularly simplex-based methods in high-dimensional settings. We propose a coordinate descent algorithm for LAD regressio

  32. Xiaoyi Li

    Post-training alignment has produced dozens of competing algorithms -- DPO, SimPO, KTO, GRPO, and others -- yet practitioners lack controlled comparisons to guide algorithm selection. We present OXRL, a unified framework implementing 51 post-training algorithms with identical infrastructure, enabling the first large-scale apples-to-apples evaluation. Our stu

  33. Matthew Brun, Xu Andy Sun, Jean-Paul Watson

    Electric power infrastructure faces increasing risk of damage and disruption due to wildfire. Operators of power grids in wildfire-prone regions must consider the potential impacts of unpredictable fires. However, traditional wildfire models do not effectively describe worst-case, or even high-impact, fire behavior. To address this issue, we propose a mixed-

  34. Yinghui Li, Jiayi Kuang, Peng Xing, Daixian Liu

    Multimodal large language models (MLLMs) perform strongly on natural images, yet their ability to understand discrete visual symbols remains unclear. We present a multi-domain benchmark spanning language, culture, mathematics, physics and chemistry, organized into three cognitive levels: perception and recognition, combination and reasoning, and association

  35. Kangyi Tian, Mingyu Xiao

    The Kidney Exchange Problem is a prominent challenge in healthcare and economics, arising in the context of organ transplantation. It has been extensively studied in artificial intelligence and optimization. In a kidney exchange, a set of donor-recipient pairs and altruistic donors are considered, with the goal of identifying a sequence of exchange -- compri

  36. Baiqiang Wang, Yan Bai, Juan Li

    The integration of Large Language Models (LLMs) into cybersecurity education for criminal justice professionals is currently hindered by the "statelessness" of reactive chatbots and the risk of hallucinations in high-stakes legal contexts. To address these limitations, we propose the CyberJustice Tutor, an educational dialogue system powered by an Agentic AI

  37. Hanyi Zhang, Wanting Ma, Wen Zhou, Xueqi Xing

    Photonic computing using chalcogenide phase-change materials (PCMs) is under active development for energy-efficient artificial intelligence (AI) applications. A key requirement is to enable as many optically programmable levels per device as possible, while maintaining relatively low optical loss. In this work, we carry out multiscale simulations using dens

  38. Quan-Yun Guo, Jing Liu, Dian-Yong Chen

    In the present work, we propose to investigate the production of $d_{N \Omega}$ in the $\Omega^{-} d \rightarrow p d_{N \Omega}^-$ process by utilizing an effective Lagrangian approach, where $d_{N \Omega}$ is identified as $N\Omega$ bound state with the binding energy $E_{b}=2.46$ MeV. Experimentally, the J-PARC hadron facility proposed to investigate the $

  39. Yuqi Yang, Dongliang Chang, Yijia Ling, Ruoyi Du

    Colour is one of the most perceptually salient yet least controllable attributes in image generation. Although recent diffusion models can modify object colours from user instructions, their results often deviate from the intended hue, especially for fine-grained and local edits. Early text-driven methods rely on discrete language descriptions that cannot ac

  40. Jiyao Liu, Junzhi Ning, Wanying Qu, Lihao Liu

    Existing medical image restoration (Med-IR) methods are typically modality-specific or degradation-specific, failing to generalize across the heterogeneous degradations encountered in clinical practice. We argue this limitation stems from the isolation of Med-IR from medical image quality assessment (Med-IQA), as restoration models without explicit quality u

  41. Wei-Wei Qi

    In this paper, we employ the Wilf-Zeilberger (WZ) method to prove a supercongruence conjecture posed by Z.-W. Sun: for any prime $p$, \begin{align*} \sum_{k=0}^{\frac{p-3}{2}}\frac{92k^2+61k+9}{(2k+1)64^k}{2k \choose k}{3k \choose k}{4k \choose 2k}\equiv 6p+16p^2\left(\frac{-1}{p}\right) \pmod{p^3}, \end{align*} where $\left(\frac{\cdot}{p}\right)$ denotes t

  42. Yan Li, Yifei Xing, Xiangyuan Lan, Xin Li

    In the era of large-scale pre-trained models, effectively adapting general knowledge to specific affective computing tasks remains a challenge, particularly regarding computational efficiency and multimodal heterogeneity. While Transformer-based methods have excelled at modeling inter-modal dependencies, their quadratic computational complexity limits their

  43. Kazuya Nishimura, Ryoma Bise, Shinnosuke Matsuo, Haruka Hirose

    Estimating slide- and patch-level gene expression profiles from pathology images enables rapid and low-cost molecular analysis with broad clinical impact. Despite strong results, existing approaches treat gene expression as a mere slide- or spot-level signal and do not incorporate the fact that the measured expression arises from the aggregation of underlyin

  44. Vahid Monfared, Mohammad Hadi Gharib, Ali Sabri, Maryam Shahali

    Prostate cancer is a leading cause of mortality in men, yet interpretation of T2-weighted prostate MRI remains challenging due to subtle and heterogeneous lesions. We developed an interpretable framework for automatic cancer detection using a small dataset of 162 T2-weighted images (102 cancer, 60 normal), addressing data scarcity through transfer learning a

  45. Xiangxu Zhang, Xiao Zhou, Hongteng Xu, Jianxun Lian

    Medication recommendations aim to generate safe and effective medication sets from health records. However, accurately recommending medications hinges on inferring a patient's latent clinical condition from sparse and noisy observations, which requires both (i) preserving the visit-level combinatorial semantics of co-occurring entities and (ii) leveraging in

  46. Haisheng Zhu, Taotao He, Mohit Tawarmalani

    We present a novel relaxation framework for general mixed-integer nonlinear programming (MINLP) grounded in computational geometry. Our approach constructs polyhedral relaxations by convexifying finite sets of strategically chosen points, iteratively refining the approximation to converge toward the simultaneous convex hull of factorable function graphs. The

  47. Jordan Hines, Corey Ostrove, Kenneth Rudinger, Stefan Seritan

    Quantum error correction (QEC), the lynchpin of fault-tolerant quantum computing (FTQC), is designed and validated against well-behaved Pauli stochastic error models. But in real-world deployment, QEC protocols encounter a vast array of other errors -- coherent and non-Pauli errors -- whose impacts on quantum circuits are vastly different than those of stoch

  48. Francesco Fournier-Facio, Rufus Willett

    We establish Kirchberg's Local Lifting Property and Lubotzky--Shalom's Property FD for classes of finitely generated groups of central importance in geometric and combinatorial group theory: $3$-manifold groups, limit groups, and certain one-relator groups and right-angled Artin groups. We deduce that such groups are very flexibly stable, with respect to nor

  49. Arushi Rai, Adriana Kovashka

    Video-LLMs often attend to irrelevant frames, which is especially detrimental for sports coaching tasks requiring precise temporal grounding. Yet obtaining frame-level supervision is challenging: expensive to collect from humans and unreliable from other models. We improve temporal grounding without additional annotations by exploiting the observation that r

  50. Jinghan Yu, Fady Alajaji, Bahman Gharesifard

    We introduce the P\'olya threshold graph model and derive its stochastic and algebraic properties. This random threshold graph is generated sequentially via a two-color P\'olya urn process. Starting from an empty graph, each time step involves a draw from the urn that produces an indicator variable, determining whether a newly added node is universal (connec

  51. David E. J. van Wijk, Dohyun Lee, Ersin Das, Tamas G. Molnar

    Guaranteeing the safety of nonlinear systems with bounded inputs remains a key challenge in safe autonomy. Backup control barrier functions (bCBFs) provide a powerful mechanism for constructing controlled invariant sets by propagating trajectories under a pre-verified backup controller to a forward invariant backup set. While effective, the standard bCBF met

  52. Yue Zhao, Yujia Gong, Ruigang Liang, Shenchen Zhu

    The widespread deployment of large language models (LLMs) calls for post-hoc methods that can flexibly adapt models to evolving safety requirements. Meanwhile, the rapidly expanding open-source LLM ecosystem has produced a diverse collection of models that already exhibit various safety-related functionalities. This motivates a shift from constructing safety

  53. Heng Ping, Peiyu Zhang, Zhenkun Wang, Shixuan Li

    Applying large language models (LLMs) to RTL code optimization for improved power, performance, and area (PPA) faces two key challenges: ensuring functional correctness of optimized designs despite LLM hallucination, and systematically prioritizing power reduction within the multi-objective PPA trade-off space. We propose POET (Power-Oriented Evolutionary Tu

  54. Haoxin Liu, Harshavardhan Kamarthi, Zhiyuan Zhao, Hongjie Chen

    Shot language understanding (SLU) is crucial for cinematic analysis but remains challenging due to its diverse cinematographic dimensions and subjective expert judgment. While vision-language models (VLMs) have shown strong ability in general visual understanding, recent studies reveal judgment discrepancies between VLMs and film experts on SLU tasks. To add

  55. Chuxuan Hu, Philip Li, Maxwell Yang, Daniel Kang

    During research, domain experts often ask analytical questions whose answers require integrating data from a wide range of web sources. Thus, they must spend substantial effort searching, extracting, and organizing raw data before analysis can begin. We formalize this process as the SODIUM task, where we conceptualize open domains such as the web as latent d

  56. Lang Zhou, Shuxuan Li, Zhuohao Li, Shi Liu

    Long-context inference remains challenging for large language models due to attention dilution and out-of-distribution degradation. Context selection mitigates this limitation by attending to a subset of key-value cache entries, yet most methods allocate a fixed context budget throughout decoding despite highly non-uniform token-level contextual demands. To

  57. Leyuan Fang, Zan Mao, Zijing Wang, Yinlong Yan

    Zero-shot object-goal navigation aims to find target objects in unseen environments using only egocentric observation. Recent methods leverage foundation models' comprehension and reasoning capabilities to enhance navigation performance. However, when faced with poor viewpoints or weak semantic cues, foundation models often fail to support reliable reasoning

  58. Victor M. Banda Guzman, James P. Edwards, C. Moctezuma Mata Zamora, Luis A. Rodriguez Chacon

    The worldline formalism allows one to obtain compact integral representations combining the information of large numbers of Feynman diagrams. However, their analytic calculation leads to a non-standard integration problem for which existing mathematical algorithms are of little help. Here I will summarize the state-of-the-art of worldline integration focusin

  59. Thierry De Pauw

    We use localized topologies to prove existence and optimal regularity results for the divergence equation $\mathrm{div} (v) = F$ in critical cases $v \in L_1(\Omega;\mathbb{R}^m)$ or $v \in C_0(\Omega;\mathbb{R}^m)$, i.e. we characterize those $F$ for which a solution $v$ exists whose norm is bounded by an appropriate norm of $F$. We assume $\Omega$ satisfie

  60. Dong Li, Zhengzhang Chen, Xujiang Zhao, Linlin Yu

    Uncovering causal structures from observational data is crucial for understanding complex systems and making informed decisions. While reinforcement learning (RL) has shown promise in identifying these structures in the form of a directed acyclic graph (DAG), existing methods often lack efficiency, making them unsuitable for online applications. In this pape

  61. Norman Guo, Wei Jiang, Yaswanth Pothuru, Baozhong Yang

    This paper provides a behavioral analysis of the post-pandemic transformation of work, using a dataset of approximately 41 billion mobile geolocation records from 73.5 million individuals in the five largest U.S. metropolitan areas from the pre- to post- pandemic periods. By tracking movements between corporate headquarters, residences, and other points of i

  62. Raghu Kulkarni

    We construct a three-dimensional Calderbank-Shor-Steane (CSS) stabilizer code on the Face-Centered Cubic (FCC) lattice. Physical qubits reside on the edges of the lattice (coordination $K=12$); X-stabilizers act on octahedral voids and Z-stabilizers on vertices, both with uniform weight 12. Computational verification confirms CSS validity ($H_{X}H_{Z}^{T}=0$

  63. Xinyi Song, Jun Yang, Yueyun Ouyang

    More than 200 moons exist in our Solar System, yet no exomoon has been confirmed to date. While the innermost two planets of the Solar System lack natural satellites and most studies favour the existence of exomoons around long-period planets, some theoretical studies that take tidal dissipation, orbital decay, and migration processes into account suggest th

  64. Wael AbdAlmageed

    Neuro-symbolic artificial intelligence (AI) systems typically couple a neural perception module to a discrete symbolic solver through a non-differentiable boundary, preventing constraint-satisfaction feedback from reaching the perception encoder during training. We introduce AS2 (Attention-Based Soft Answer Sets), a fully differentiable neuro-symbolic archit

  65. Richard Montgomery

    We pose several questions for the classical N-body problem inspired by connections between the virial equation and the Jacobi-Maupertuis formulationof mechanics. We answer some.

  66. Md Takrim Ul Alam, Akif Islam, Mohd Ruhul Ameen, Abu Saleh Musa Miah

    Large language models (LLMs) deployed behind APIs and retrieval-augmented generation (RAG) stacks are vulnerable to prompt injection attacks that may override system policies, subvert intended behavior, and induce unsafe outputs. Existing defenses often treat prompts as flat strings and rely on ad hoc filtering or static jailbreak detection. This paper propo

  67. Runze Yang, Longbing Cao, Xiaoming Wu, Xin You

    Separating multiple effects in time series is fundamental yet challenging for time-series forecasting (TSF). However, existing TSF models cannot effectively learn interpretable multi-effect decomposition by their smoothing-based temporal techniques. Here, a new interpretable frequency-based decomposition pipeline MLOW captures the insight: a time series can

  68. Xiaoxu Ma, Dong Li, Minglai Shao, Xintao Wu

    Text-attributed graphs, where nodes are enriched with textual attributes, have become a powerful tool for modeling real-world networks such as citation, social, and transaction networks. However, existing methods for learning from these graphs often assume that the distributions of training and testing data are consistent. This assumption leads to significan

  69. Zhuoyue Chen, Kechao Cai

    Quantum multi-armed bandits (MAB) and stochastic linear bandits (SLB) have recently attracted significant attention, as their quantum counterparts can achieve quadratic speedups over classical MAB and SLB. However, most existing quantum MAB algorithms assume ideal quantum Monte Carlo (QMC) procedures on noise-free circuits, overlooking the impact of noise in

  70. Yibo Shi, Jungang Li, Linghao Zhang, Zihao Dongfang

    Long-horizon GUI agents are a key step toward real-world deployment, yet effective interaction memory under prevailing paradigms remains under-explored. Replaying full interaction sequences is redundant and amplifies noise, while summaries often erase dependency-critical information and traceability. We present AndroTMem, a diagnostic framework for anchored

  71. Asmita Bhardwaj, Yuya Jeremy Ong, Eelaaf Zahid, Basel Shbita

    Decoding strategies largely determine the quality of Large Language Model (LLM) outputs, yet widely used heuristics such as greedy or fixed temperature/top-p decoding are static and often task-agnostic, leading to suboptimal or inconsistent generation quality across domains that demand stylistic or structural flexibility. We introduce a reinforcement learnin

  72. Huy Che, Dinh-Duy Phan, Duc-Khai Lam

    Collecting and annotating datasets for pixel-level semantic segmentation tasks are highly labor-intensive. Data augmentation provides a viable solution by enhancing model generalization without additional real-world data collection. Traditional augmentation techniques, such as translation, scaling, and color transformations, create geometric variations but f

  73. Minjun Kim, Jaehyeon Choi, Hyunwoo Yang, Jongjin Kim

    What happens when multiple compression methods are combined-does the order in which they are applied matter? Joint model compression has emerged as a powerful strategy to achieve higher efficiency by combining multiple methods such as pruning and quantization. A central but underexplored factor in joint model compression is the compression order, or the sequ

  74. Anay Agarwalla, Simeon Sayer

    Sensitive attributes are legally protected characteristics that should not be used to discriminate. Careful steps have been taken to minimize the risk of human bias regarding these fields, such as race and age. Large language models (LLMs) are similarly trained not to attempt to infer these aspects. However, just because they shouldn't, doesn't mean they don

  75. Kaan T. Gun, Xiaozhe Wang, Danial Jafarigiv

    Electric vehicles (EVs) in Vehicle-to-Grid (V2G) systems act as distributed energy resources that support grid stability. Centralized coordination such as the extended State Space Model (eSSM) enhances scalability and estimation efficiency but may introduce new cyber-attack surfaces. This paper presents a stealthy False Data Injection Attack (FDIA) targeting

  76. Minjun Kim, Jongjin Kim, U Kang

    How can we accurately quantize a pre-trained model without any data? Quantization algorithms are widely used for deploying neural networks on resource-constrained edge devices. Zero-shot Quantization (ZSQ) addresses the crucial and practical scenario where training data are inaccessible for privacy or security reasons. However, three significant challenges h

  77. Massimiliano de Sa, Aaron D. Ames

    In 1983, Brockett developed a topological necessary condition for the existence of continuous, asymptotically stabilizing control laws. Building upon recent work on necessary conditions for set stabilization, we develop Brockett-like necessary conditions for the existence of control barrier functions (CBFs). By leveraging the unique geometry of CBF safe sets

  78. Zhanjie Wen, Wenxiu Li, Jiechang Xia, Jingqiao Guo

    In the context of the rapid development of digital finance, some financial technology companies exhibit the phenomenon of "AI washing," where they overstate their AI capabilities while underinvesting in actual AI resources. This paper constructs a corporate-level AI washing index based on CHFS2019 data and AI investment data from 15-20 financial technology c

  79. Jason Dury

    Embedding models group text by semantic content, what text is about. We show that temporal co-occurrence within texts discovers a different kind of structure: recurrent transition-structure concepts or what text does. We train a 29.4M-parameter contrastive model on 373 million co-occurrence pairs from 9,766 Project Gutenberg texts (24.96 million passages), m

  80. Axel Guinot, Rachel Mandelbaum

    Weak gravitational lensing is a widely used probe in cosmological analysis. It allows astrophysists to understand the content and evolution of the Universe. We are entering an era where we are not limited by the data volume but by systematic uncertainties. It is in this context that we present here a simple python-based software package to help in the comput

  81. Yang Liu, Jiyao Yang, Hongjin Zhao, Xiaoyong Li

    Large vision-language models (LVLMs) demonstrate strong performance in dermatology; however, evaluating diagnostic reasoning for rare conditions remains largely unexplored. Existing benchmarks focus on common diseases and assess only final accuracy, overlooking the clinical reasoning process, which is critical for complex cases. We address this gap by constr

  82. Arundhathi Dev, Justin Zhan

    Sparse attention mechanisms promise to break the quadratic bottleneck of long-context transformers, yet production adoption remains limited by a critical usability gap: optimal hyperparameters vary substantially across layers and models, and current methods (e.g., SpargeAttn) rely on manual grid search to identify them. We propose AFBS-BO (Adaptive Fidelity

  83. Lehel Csillag, Nicoleta Voicu, Salah Elgendi, Christian Pfeifer

    In metric-affine geometry, autoparallels are generically non-variational, i.e., they are not the extremals of any action integral. The existence of a parametrization-invariant action principle for autoparallels is a long-standing open problem, which is equivalent to the so-called Finsler metrizability of the connection -- that is, to the fact that these auto

  84. Li Wenxiu, Wen Zhanjie, Xia Jiechang, Guo Jingqiao

    At a time when the phenomenon of 'AI washing' is quietly spreading, an increasing number of enterprises are using the label of artificial intelligence merely as a cosmetic embellishment in their annual reports, rather than as a genuine engine driving transformation. A test regarding the essence of innovation and the authenticity of information disclosure has

  85. Yu-Zhuo Li, Li-Chao Peng, Ke-Mi Xu

    Negativities in quasiprobability distributions, a foundational concept originating in quantum optics, serve as a fundamental signature of quantum nonclassicality, with entanglement quasiprobabilities offering a necessary and sufficient criterion for entanglement. However, practical reconstruction of entanglement quasiprobabilities conventionally requires ful

  86. Yugo Miyata, Tomohiro Shiraishi, Shuichi Nishino, Ichiro Takeuchi

    A data analysis pipeline is a structured sequence of steps that transforms raw data into meaningful insights by integrating multiple analysis algorithms. In many practical applications, analytical findings are obtained only after data pass through several data-dependent procedures within such pipelines. In this study, we address the problem of quantifying th

  87. Koichi Hamaguchi, Atsuya Niki, Kwok Hei To

    Despite being a simple and well-motivated thermal relic scenario, coannihilation dark matter (DM) has remained largely unexplored experimentally due to the difficulty of probing its nearly degenerate mass spectrum. Recent LHC searches, however, have significantly improved the sensitivity to such compressed spectra, motivating a reassessment of the viable par

  88. Arushi Rai, Qiang Zhang, Hanqing Zeng, Yunkai Zhang

    Large language models (LLMs) exhibit strong reasoning capabilities but typically require expensive post-training to reach high performance. Recent test-time alignment methods offer a lightweight alternative, but have been explored mainly for preference alignment rather than reasoning. To bridge this gap, we propose, Token-level Adaptive Routing (TARo), which

  89. Hanwen Wang, Zhenlong Fang, Josiah Hanna, Xiaobin Xiong

    In this paper, we present a hardware-control co-design approach that enables efficient and versatile roller skating on quadrupedal robots equipped with passive wheels. Passive-wheel skating reduces leg inertia and improves energy efficiency, particularly at high speeds. However, the absence of direct wheel actuation tightly couples mechanical design and cont

  90. Janani S K, Kushagra Gupta, Ufuk Topcu, David Fridovich-Keil

    A fundamental problem in noncooperative dynamic game theory is the computation of Nash equilibria under different information structures, which specify the information available to each agent during decision-making. Prior work has extensively studied equilibrium solutions for two canonical information structures: feedback, where agents observe the current st

  91. Songfeng Zhu

    In incremental classification tasks for hyperspectral images, catastrophic forgetting is an unavoidable challenge. While memory recall methods can mitigate this issue, they heavily rely on samples from old categories. This paper proposes a teacher-based knowledge retention method for incremental image classification. It alleviates model forgetting of old cat

  92. Joshua Cesca, Archil Kobakhidze

    We study the cosmological implications of the minimal non-linear realisation of scale invariance within the Standard Model (SM). This framework provides a technically natural explanation for the hierarchy between the Planck scale and the electroweak scale and introduces only a light, feebly coupled dilaton field beyond the SM particles. Although the model is

  93. Zameddin I. Ismailov, Sergei Silvestrov, Pembe Ipek Al

    In this study, the classical results on the joint numerical radius for $n$-tuples of Hilbert space operators are extended to the setting of the joint $(f,\delta)$-numerical radius. New and diverse contributions to this area are provided, including novel estimates for the lower and upper bounds of the $(f,\delta)$-numerical radius in the context of sectorial

  94. Bohan Wu, Julius von Kügelgen, David M. Blei

    Causal representation learning (CRL) aims to learn low-dimensional causal latent variables from high-dimensional observations. While identifiability has been extensively studied for CRL, estimation has been less explored. In this paper, we explore the use of empirical Bayes (EB) to estimate causal representations. In particular, we consider the problem of le

  95. Changxiao Nigel Shen, Wim M. van Rees

    Wavelet-based grid adaptation methods use multiresolution analysis for error estimation, offering a mathematically rigorous approach to adaptive grid refinement when solving Partial Differential Equations (PDEs). However, applying these methods to PDE discretizations with immersed geometries is challenging, as standard interpolating wavelet transforms lose c

  96. Yonghan Lee, Dinesh Manocha

    We present Inst4DGS, an instance-decomposed 4D Gaussian Splatting (4DGS) approach with long-horizon per-Gaussian trajectories. While dynamic 4DGS has advanced rapidly, instance-decomposed 4DGS remains underexplored, largely due to the difficulty of associating inconsistent instance labels across independently segmented multi-view videos. We address this chal

  97. Zeeshan Ahmad, Naimul Khan

    Vision Transformers have shown tremendous success in numerous computer vision applications; however, they have not been exploited for stress assessment using physiological signals such as Electrocardiogram (ECG). In order to get the maximum benefit from the vision transformer for multilevel stress assessment, in this paper, we transform the raw ECG data into

  98. Oleksii Nasypanyi, Francois Rameau

    Keypoint matching can be slow and unreliable in challenging conditions such as repetitive textures or wide-baseline views. In such cases, known geometric relations (e.g., the fundamental matrix) can be used to restrict potential correspondences to a narrow epipolar envelope, thereby reducing the search space and improving robustness. These epipolar-guided ma

  99. Anastasios Manganaris, Jeremy Lu, Ahmed H. Qureshi, Suresh Jagannathan

    Sequences of interdependent geometric constraints are central to many multi-agent Task and Motion Planning (TAMP) problems. However, existing methods for handling such constraint sequences struggle with partially ordered tasks and dynamic agent assignments. They typically assume static assignments and cannot adapt when disturbances alter task allocations. To

  100. Kaijie Xu, Yiwei Zhang, Brian Yang, Clark Verbrugge

    Open-world missions often rely on repeated formulas, yet designers lack systematic ways to examine pacing, variation, and experiential balance across large portfolios. We introduce the Mission Action Quality Vector (MAQV), a six-dimensional framework-covering combat, exploration, narrative, emotion, problem-solving, and uniqueness-paired with an action block