Skip to content

April 2026 arXiv papers — page 100

Showing 9,90110,000 of 25,062 papers

  1. Diego Forlivesi, Lorenzo Valentini, Marco Chiani

    Reliable quantum computation requires fault-tolerant protocols to prevent errors from propagating during syndrome extraction in quantum error correction. We present a novel fault-tolerant syndrome extraction technique for CSS codes, which we refer to as the cut-cat state scheme. While each ancilla qubit interacts non-fault-tolerantly with a pair of data qubi

  2. Jingbo Sun, Wenyue Chong, Songjun Tu, Qichao Zhang

    Agentic retrieval-augmented generation (RAG) systems enable large language models (LLMs) to solve complex tasks through multi-step interaction with external retrieval tools. However, such multi-step interaction often involves redundant search steps, incurring substantial computational cost and latency. Prior work limits search depth (i.e., the number of sear

  3. Vilde G. S. Lunde, Øystein S. Fjellvåg, Allan M. Döring, Marc Straßheim

    We have investigated NdCo2-xNix cubic Laves compounds with 0 <= x <= 1 using neutron diffraction and bulk magnetization measurements to study the influence of partial Ni substitutions of Co on the phase transitions and the magnetocaloric effect. Upon cooling, NdCo2 undergoes a cubic to tetragonal transition at 100 K, and a tetragonal to orthorhombic transiti

  4. Yan Guo, Yanjin Wang

    Inflow BC plays a critical role in the study of hyperbolic PDE in a bounded domain. We establish $W^{1,\infty}$ stability for 1D hyperbolic conservation laws with inflow data in a bounded interval, and $W^{2,3+}$ stability of a large class of shear flows for the 3D incompressible Euler system with inflow BC in finite square or circular pipes.

  5. Mengju Yuan, Yugang Zhang, Ying Zhu, Jingchun Gao

    The dynamics of phase boundaries, such as superconducting/normal (S/N) interfaces in type-I superconductors, are typically obscured in conventional magnetic measurements, which are dominated by surface barriers and over-damped flux processes. Here, we employ ac magnetostriction as a sensitive probe to reveal the distinct bulk dynamics of these domain walls i

  6. Fumio Ishizaki

    We study stochastic processes on combinatorial state spaces with local transition constraints, as arise in local search algorithms. We show that asymmetry in local transitions induces a systematic drift in a distance process relative to a reference configuration. This drift results from the imbalance between inward and outward transitions, translating combin

  7. Filip Chudy, Paweł Woźny

    We present new representations of Gauss--Legendre polynomials and their derivatives in the shifted power basis and in bases related to symmetric orthogonal Jacobi polynomials. Using these representations and certain recurrence relations, we propose efficient $O(n^2+dn)$ methods for evaluating a Gauss--Legendre curve of degree $n$ in $\mathbb E^d$. We also pr

  8. Juraj Lorincik, Hannah Collier, Vanessa Polito, Laura A. Hayes

    Apparent slipping motions of flare ribbon kernels and the formation of hard X-ray (HXR) footpoints are important signatures of magnetic reconnection in solar flares. Ultraviolet (UV) and HXR ribbon emission can show quasi-periodic pulsations (QPPs), but the link between HXR QPP sources and slipping reconnection remains poorly understood. In this work, we ana

  9. Oishik Chowdhury, Bastin Tony Roy Savarimuthu, Sherlock A. Licorish

    Design patterns provide reusable solutions to recurring software design problems. Automatically detecting these patterns in source code can help bootstrap new developers' understanding of unfamiliar software system architectures, and can help experienced developers to quickly identify and rectify potential quality issues. While many prior research works have

  10. George Fatouros, Kostas Metaxas

    We present the first portfolio-level validation of MarketSenseAI, a deployed multi-agent LLM equity system. All signals are generated live at each observation date, eliminating look-ahead bias. The system routes four specialist agents (News, Fundamentals, Dynamics, and Macro) through a synthesis agent that issues a monthly equity thesis and recommendation fo

  11. Xiangyu Ge, Jiafei Ge, Shengmei Zhao, Le Wang

    Quantum Noise Characterization (QNC) is indispensable for benchmarking and mitigating errors in Noisy Intermediate-Scale Quantum (NISQ) devices. However, traditional Quantum Process Tomography (QPT) suffers from an exponential parameter explosion, severely hindering its scalability. In this paper, we propose a Hierarchical Progressive Optimization (HPO) fram

  12. Jiaang Li, Zhendong Mao, Quan Wang, Yuning Wan

    Retrieval-Augmented Generation (RAG) enhances the factuality of Large Language Models (LLMs) by incorporating retrieved documents and/or generated context. However, LLMs often exhibit a stylistic bias when presented with mixed contexts, favoring fluent but hallucinated generated content over factually grounded yet disorganized retrieved evidence. This phenom

  13. Kyeongman Park, Minha Jhang, Kyomin Jung

    Modern generative models still lack human-level creativity, particularly in multi-branch diversity. Prior approaches to address this problem often incur heavy computation or strong dependency on model architecture. Therefore, we introduce UAG(Universal Avoidance Generation), a model-agnostic and computationally efficient generation strategy that penalizes si

  14. Pradipta Biswas, Himanshu Vishwakarma, Mukund Mitra, KamalPreet Singh Saluja

    Human Space Flight missions often require interaction with touchscreen displays. This paper presents a study of investigating human machine interaction with touchscreen using both finger and stylus in the International Space Station. The study also reports cognitive state of astronauts in the form of spatial 2-back test and mental well-being through self-rep

  15. Raghavendra Ramachandra

    Face morphing attacks pose a substantial risk to the reliability of face recognition systems used in passport issuance, border control, and digital identity verification. Detecting morphing attacks from a single facial image remains challenging owing to the lack of a trusted reference and the diversity of attack generation methods. This paper presents a new

  16. Xinqing Li, Xin He, Xindong Zhang, Ming-Ming Cheng

    Deploying Vision-Language Models (VLMs) under aggressive low-bit inference remains challenging because inference cost is dominated by the long visual-token prefix during prefill and the growing KV cache during autoregressive decoding. Token pruning and low-bit quantization are complementary for reducing these costs, yet naive stage-wise combinations are ofte

  17. Meng Zhang, Jinzhong Ning, Xiaolong Wu, Hongfei Lin

    Grounded Multimodal Named Entity Recognition (GMNER) aims to jointly identify named entity mentions in text, predict their semantic types, and ground each entity to a corresponding visual region in an associated image. Existing approaches predominantly adopt pipeline-based architectures that decouple textual entity recognition and visual grounding, leading t

  18. Akash Ghosh, Subhadip Baidya, Sriparna Saha, Xiuying Chen

    Vision-Language Models (VLMs) are increasingly used in clinical diagnostics, yet their robustness to adversarial attacks remains largely unexplored, posing serious risks. Existing medical attacks focus on secondary objectives such as model stealing or adversarial fine-tuning, while transferable attacks from natural images introduce visible distortions that c

  19. Alberto Testoni, Iacer Calixto

    Safe clinical deployment of Large Language Models (LLMs) requires not only high accuracy but also robust uncertainty calibration to ensure models defer to clinicians when appropriate. Our paper investigates how social descriptors of a patient (specifically sexual orientation and religious affiliation) distort these uncertainty signals and model accuracy. Eva

  20. Sudipto Chowdhury, Shallu

    This article examines the Dirichlet boundary control problem governed by the Poisson equation, where the control variables are square integrable functions defined on the boundary of a two-dimensional bounded, convex, polygonal domain. It employs an ultra-weak formulation and utilizes Crouzeix-Raviart finite elements to discretize the state variable, while em

  21. Linjie Ma

    In high-contrast composites, the electric (or stress) field may exhibit significant amplification in the narrow region between inclusions. The behavior of the solution depends on the distance $\epsilon$ between the inclusions, which tends to $0$. The purpose of this paper is to provide a simple proof of optimal pointwise estimates for the insulated conductiv

  22. Rina Mishra, Gaurav Varshney, Doddipatla Sesha Sahithi

    The rapid adoption of open-source Large Language Models (LLMs) in offline and enterprise environments has introduced a largely unexamined security risk like susceptibility to adversarial phishing prompts under static safety configurations. In this work, we systematically investigate this vulnerability through GuardPhish, a large scale multi-vector phishing p

  23. Zhiyin Yu, Bo Zhang, Qibin Hou, Zhonghai Wu

    Previous LLMs-based RL studies typically follow either supervised learning with high annotation costs, or unsupervised paradigms using voting or entropy-based rewards. However, their performance remains far from satisfactory due to the substantial annotation cost and issues such as model collapse or reward hacking. To address these issues, we introduce a new

  24. Zhiyin Yu, Yuchen Mou, Juncheng Yan, Junyu Luo

    Reinforcement learning (RL) has emerged as a powerful post-training paradigm for enhancing the reasoning capabilities of large language models (LLMs). However, reinforcement learning for LLMs faces substantial data scarcity challenges, including the limited availability of high-quality external supervision and the constrained volume of model-generated experi

  25. Zihao Ren, Lei Wang, Guodong Shi

    Various distributed gradient descent algorithms for multi-agent optimization have incorporated the Nesterov accelerated gradient method, where the use of momentum enhances convergence rates. These algorithms have found broad applications in large-scale machine learning and optimization owing to their simplicity and low communication complexity. In this paper

  26. Marcel Kollovieh, Sirine Ayadi, Stephan Günnemann

    Discrete diffusion models form a powerful class of generative models across diverse domains, including text and graphs. However, existing approaches face fundamental limitations. Masked diffusion models suffer from irreversible errors due to early unmasking, while uniform diffusion models, despite enabling self-correction, often yield low-quality samples due

  27. Guangsheng Yu, Xu Wang

    Research artifacts are distributed primarily as reader-oriented documents like PDFs. This creates a bottleneck for increasingly agent-assisted and agent-native research workflows, in which LLM agents need to infer fine-grained, task-relevant information from lengthy full documents, a process that is expensive, repetitive, and unstable at scale. We introduce

  28. Ziao Zhang, Kou Shi, Shiting Huang, Avery Nie

    As the capability frontier of autonomous agents continues to expand, they are increasingly able to complete specialized tasks through plug-and-play external skills. Yet current benchmarks mostly test whether models can use provided skills, leaving open whether they can discover skills from experience, repair them after failure, and maintain a coherent librar

  29. Enrui Yang, Yuezun Li

    Detecting face forgeries using CLIP has recently emerged as a promising and increasingly popular research direction. Owing to its rich visual knowledge acquired through large-scale pretraining, most existing methods typically rely on the visual encoder of CLIP, while paying limited attention to the text modality. Given the instructive nature of the text moda

  30. Jiatong Li, Zheng Chen, Kai Liu, Jingkai Wang

    This paper provides a review of the NTIRE 2026 challenge on mobile real-world image super-resolution, highlighting the proposed solutions and the resulting outcomes. The challenge aims to recover high-resolution (HR) images from low-resolution (LR) counterparts generated through unknown degradations with a x4 scaling factor while ensuring the models remain e

  31. Jianing Hao, Yuhe Wu, Yuanjian Xu, Shichang Meng

    Large language models (LLMs) hold great promise for business applications, yet business analysis remains inherently complex, demanding rigorous reasoning and the integration of diverse knowledge sources. Existing benchmarks typically target narrow tasks and thus leave a fundamental question unanswered: how can LLMs be reliably applied in business, and how ar

  32. Jiakun Li, Xingwei He, Kefan Li, Hongzheng Chai

    Test-time scaling improves the reasoning performance of large language models but often results in token-inefficient overthinking, where models continue reasoning beyond what is necessary for a correct answer. Existing dynamic early-exit methods typically rely on single-step confidence signals, which are often unreliable for detecting reasoning convergence i

  33. Sooraj M, Moumanti Podder, Archi Roy

    We study a model of market economics wherein the $(n+1)$-st customer, for each $n\geqslant N$, with $N$ being a prespecified positive integer, draws a sample of (random) size $K_{n}$, either with replacement or without, from the customers of the past. Each sampled customer is queried as to which of the two products, A and B, available in the oligopolistic ma

  34. Chinthakuntla Meghan Sai, Murarisetty V Sai Kartheek, Sita Devi Bharatula, Karthik Seemakurthy

    The scarcity of labeled clinical data in oncology makes Few-Shot Learning (FSL) a critical framework for Computer Aided Diagnostics, but we observed that standard Prototypical Networks often struggle with the "prototype instability" caused by morphological noise and high intra-class variance in brain tumor scans. Our work attempts to minimize this by integra

  35. Tiankai Yang, Yi Nian, Xinyuan Li, Ruiyao Xu

    Aligning large language models with human preferences must balance two competing goals: responding helpfully to legitimate requests and reliably refusing harmful ones. Most preference-based safety alignment methods collapse safety into a single scalar that is applied uniformly to every preference pair. The result is a model that looks safe on average but sta

  36. Chenxing Li, Yiping Duan, Xiaoming Tao

    Video Scene Graph Generation aims to obtain structured semantic representations of objects and their relationships in videos for high-level understanding. However, existing methods still have limitations in handling long-tail distributions. This paper proposes the Frequency-guided Relational Multi-level Reasoning (FReMuRe) model, which enhances the modeling

  37. Yangsong Lan, Hongliang Dai, Piji Li

    Long Chain-of-Thought (CoT) reasoning is pivotal for the success of recent reasoning models but suffers from high computational overhead and latency. While prior works attempt to compress CoT via external compressor, they often fail to align with the model's internal reasoning dynamics, resulting in the loss of critical logical steps. This paper presents \te

  38. Øystein Linnebo

    Potentialism is the view that objects are successively generated in an incompletable process. A strict version of the view adds that truths are successively determined. Strict potentialism can be analyzed using two modalities: one for the generation of objects, another for truths becoming determined. The result is a classical bimodal logic. We obtain simpler

  39. Yueyang Ding, HaoPeng Zhang, Rui Dai, Yi Wang

    Comprehensive understanding of time series remains a significant challenge for Large Language Models (LLMs). Current research is hindered by fragmented task definitions and benchmarks with inherent ambiguities, precluding rigorous evaluation and the development of unified Time Series Reasoning Models(TSRMs). To bridge this gap, we formalize Time Series Reaso

  40. Khachatur A. Khachatryan

    We introduce and study a new class of nonlinear monotone operators acting in normal cones of real Banach spaces and possessing the property of strong concavity. We establish new constructive principles for the existence of nonzero fixed points for this class of operators. Further, we prove that the corresponding iterative process converges to the fixed point

  41. Jingyi Ren, Ante Wang, Yunghwei Lai, Xiaolong Wang

    Reliable Large Language Models (LLMs) should abstain when confidence is insufficient. However, prior studies often treat refusal as a generic "I don't know'', failing to distinguish input-level ambiguity (data uncertainty) from capability limitations (model uncertainty). This lack of distinction limits downstream action decisions like requesting clarificatio

  42. Mohamed Boucetta

    In this paper, we establish a complete structural description of flat Lorentzian Lie groups, i.e., Lie groups endowed with a flat left invariant Lorentzian metric, thereby resolving a long-standing open problem in the theory of pseudo-Riemannian Lie groups. Our main result shows that any flat Lorentzian Lie group either admits a timelike parallel left-invari

  43. Moo K. Chung, Luigi Maccotta, Aaron Struck

    Cortical folding reflects coordinated neurodevelopmental processes and provides a sensitive marker of neurological disease. In juvenile myoclonic epilepsy (JME), structural abnormalities are subtle and spatially distributed, limiting the sensitivity of conventional morphometric measures such as cortical thickness. We introduce a Poisson flow model derived fr

  44. Zekun Xi, Yichen Nie, Ziyan Jiang, Yujie Bao

    Despite the rapid advancement of generative agents, their deployment in real-world industry scenarios often encounters significant challenges due to a lack of domain-specific knowledge. To address this gap, we present KnowPilot: a Domain-Specific Knowledge Augmented Generative Agent System. KnowPilot is an open-source framework that integrates task-specific

  45. Poorva Garg, Renato Lui Geh, Daniel Israel, Todd Millstein

    LLMs are widely used for code generation and mathematical reasoning tasks where they are required to generate structured output. They either need to reason about code, generate code for a given specification, or reason using programs of thought. The typical approach to code generation is to prompt the model and generate samples until an appropriate program i

  46. Zizhang Luo, Yansong Xu, Runlin Guo, Fan Cui

    RTL program repair remains a critical bottleneck in hardware design and verification. Traditional automatic program repair (APR) methods rely on predefined templates and synthesis, limiting their bug coverage. Large language models (LLMs) and coding agents based on them offer flexibility but suffer from randomness and context corruption when handling long RT

  47. H. M. Shadman Tabib, Tasriad Ahmed Tias, Nafis Tahmid

    Copy-move forgery, where a region within an image is duplicated to hide or fabricate content, remains a persistent threat to visual media integrity. We introduce GraphSpecForge, a training-free framework that detects copy-move forgery by analysing the spectral structure of attention graphs from a pretrained Stable Diffusion U-Net. Our central insight is that

  48. Chunliang Li, Tianze Cao, Sanyuan Zhao

    Visual Autoregressive (VAR) modeling inefficiently applies a fixed computational depth to each position when generating high-resolution images. While existing methods accelerate inference by pruning tokens using frequency maps, their binary hard-pruning approach is fundamentally limited and fails to improve quality even with better frequency estimation. Obse

  49. Johannes Bund, Amir Leshem, Moti Medina

    Metastability is a spurious mode of operation in digital signals, where an electrical signal fails to settle into a stable state within a specified time, leading to uncertainty and potentially failing downstream hardware. A system that computes the closure over all possibilities, given an uncertain input, is called a Metastability-containing system. While pr

  50. Chao Jin, Wenkui Yang, Hao Sun, Yuqi Liao

    While progress in GUI agents has been largely driven by industrial-scale training, ungrounded hallucinations often trigger cascading failures in real-world deployments.Unlike general VLM domains, the GUI agent field lacks a hallucination-focused suite for fine-grained diagnosis, reliable evaluation, and targeted mitigation.To bridge this gap, we introduce Ha

  51. Shuyue Stella Li, Bhargavi Paranjape, Kerem Oktar, Zhongyao Ma

    User preferences evolve across months of interaction, and tracking them requires inferring when a stated preference has been changed by a subsequent life event. We define this problem as long-horizon personalization and observe that progress on it is limited by data availability and measurement, with no existing resource providing both naturalistic long-hori

  52. Lingyan Wu, Xiang Zheng, Weiqi Zhai, Wei Wang

    Process-Level Reward Models (PRMs) are essential for guiding complex reasoning in large language models, yet existing PRM benchmarks cover only general domains such as mathematics, failing to address medical reasoning -- which is uniquely characterized by safety criticality, knowledge intensity, and diverse error patterns. Without a reliable medical PRM eval

  53. Kangkang Sun, Junyi He, Juntong Liu, Xiuzhen Chen

    Autonomous platoons traversing infrastructure gaps increasingly depend on LEO satellite backhaul for safety-critical updates, yet no existing framework jointly addresses compound Doppler from simultaneous satellite and vehicle motion, sub-slot handover outages that exceed collision-alert deadlines, and heterogeneous freshness requirements across three vehicu

  54. Jinzi Bai, Fei Fang

    In this paper, we study the Fu\v{c}ik spectrum for the operator with rapidly increasing weight, which is defined as a set $\Sigma$ comprising those $(\alpha, \beta) \in \mathbb{R}^2$ such that \begin{equation*} \left\{\begin{array}{l} L u:=-\Delta u-\frac{1}{2}(x \cdot \nabla u)=\alpha u^{+}-\beta u^{-}, \text{in}\ \mathbb{R}^N,\\ u\in X, \end{array}\right.

  55. Hiroshi Kozaki, Satsuki Matsuno, Tatsuhiko Koike, Yoshiyuki Morisawa

    We establish a deformation framework for highly symmetric solutions to the Einstein equations. In this framework, four-dimensional metrics are constructed from three-dimensional {\eta}-Einstein metrics admitting a deformation determined by a single function. Under this deformation, the resulting spacetime solves the Einstein equations with a string-cloud sou

  56. Xueheng Li, Tao Hu, Ke Cao, Runsheng Qi

    Effective pest recognition and management are crucial for sustainable agricultural development. However, collecting pest data in real scenarios is often challenging. Compared to other domains, pests exhibit a wide variety of species with complex and diverse morphological characteristics. Existing techniques struggle to effectively model the key visual and hi

  57. Zixin Zhou, Tianxi Jiang, Menglong Yang, Zhihua Feng

    Physical neural networks offer a transformative route to edge intelligence, providing superior inference speed and energy efficiency compared to conventional digital architectures. However, realizing scalable, end-to-end, fully analog recurrent neural networks for temporal information processing remains challenging due to the difficulty of faithfully mapping

  58. Xinxin Li, Yudong Wei, Hao Zhang

    We study the two-set feasibility problem of finding a point in the intersection $X\cap Y$ of closed convex sets in a Hilbert space. We propose a generalized composed alternating relaxed projection algorithm (gCARPA) that blends Douglas-Rachford-type and projection-reflection-type dynamics via an outer averaging step $\mu$ and an internal relaxation $(\gamma,

  59. Xiakun Li, Hao Wu, Bican Xia, Tengshun Yang

    Stochastic constraints, which incorporate both deterministic parameters and random variables, extend classical deterministic constraints by explicitly accounting for uncertainty. These constraints are increasingly prevalent in data science, artificial intelligence, and bioinformatics; however, solving them requires addressing quantitative satisfaction proble

  60. Yunkai Dang, Yifan Jiang, Yizhu Jiang, Anqi Chen

    Multimodal Large Language Models (MLLMs) have demonstrated exceptional capabilities in various perception and reasoning tasks. Despite this success, ensuring their reliability in practical deployment necessitates robust confidence estimation. Prior works have predominantly focused on text-only LLMs, often relying on computationally expensive self-consistency

  61. Samuel Sameer Tanguturi

    The most important architectural problem in AI is not the size of the model but the absence of a layer that carries forward what the model has come to understand. Sessions end. Context windows fill. Memory APIs return flat facts that the model has to reinterpret from scratch on every read. The result is intelligence that is powerful per session and amnesiac

  62. Andrea Bevilacqua, Alice Boldrin

    In this paper we introduce a new general framework for the study of phenomenological quantum gravity theories (PQG). The key idea is the introduction of two different types of spacetime, an observer-independent spacetime (modeled by a smooth orientable manifold) and an observer-dependent one (which has an inherently discrete causal set structure). The intera

  63. Ziqing Wang, Kaize Ding

    Node classification on text-attributed graphs (TAGs) is a fundamental task with broad applications in citation analysis, social networks, and recommendation systems. Current GNN-based approaches suffer from shallow text encoding and heavy dependence on labeled data, limiting their effectiveness in label-scarce settings. While large language models (LLMs) nat

  64. Aryansh Saxena, Suresh C. Jaryal, K. K. Sharma

    In this paper, we study the quasinormal modes of the generalized Joshi-Malafarina-Narayan (JMN) naked singularity spacetime using the exact Wentzel-Kramers-Brillouin (WKB) method. Working in the complex radial plane, we construct the exact WKB momentum function, determine its turning points, and compute the associated Stokes geometry for representative quasi

  65. Wenwei Xie, Jie Yin, Lu Ma, Xuansong Zhang

    AI-generated imagery has reached near-photorealistic fidelity, yet this technology poses significant threats to information security and societal trust. Existing deepfake detection methods often exhibit limited robustness in open-world scenarios. To address this limitation, this paper investigates intrinsic discrepancies between synthetic and authentic image

  66. Zhang Zhang, Yifeng Zeng, Houshi Jiang, Yinghui Pan

    WeatherSeg, an advanced semi-supervised segmentation framework, addresses autonomous driving's environmental perception challenges in adverse weather while reducing annotation costs. This framework integrates a Dual Teacher-Student Weight-Sharing Model (DTSWSM) that enables knowledge distillation from weather-affected images, and a Classifier Weight Updating

  67. Stavros Mouslopoulos

    We derive strict quantitative conditions under which a collective quantum system of N~370 spins exhibits classically forbidden temporal correlations in a spinor Bose-Einstein condensate (BEC). The Lipkin-Meshkov-Glick (LMG) model near its Z_2-breaking quantum critical point supports a mesoscopic superposition a|P> + b|R> of two macroscopic ordered phases (|P

  68. Yuxuan Yu, Jiashuo Liu, Hua Tong, Honghua Lou

    Polycube structures provide parametric domains for all-hexahedral (all-hex) mesh generation and analysis-suitable volumetric spline construction in isogeometric analysis (IGA). Recent learning-based polycube pipelines have improved automation, yet several challenges remain when handling complex CAD geometries. These challenges include the limited diversity o

  69. Sheng Zhang, Junyi Li, Yingyi Zhang, Pengyue Jia

    Recent advances in large language models (LLMs) have scaled the potential for reasoning and agentic search, wherein models autonomously plan, retrieve, and reason over external knowledge to answer complex queries. However, the iterative think-search loop accumulates long system memories, leading to memory dilution problem. In addition, existing memory manage

  70. Hongkan Chen, Qingshan Zhou, Robin Haunschild, Yi Bu

    In modern scientific collaboration networks, certain researchers play a pivotal role in bridging scholars who have never worked together - a phenomenon we term academic "match-makers." Despite their potential importance, the prevalence, characteristics, benefits, and long-term trajectory of these individuals remain underexplored. Using the Microsoft Academic

  71. Hongyu Wang, Xudong Gao, Chuanneng Luo, Haijiang Yu

    All-optical switching technology is a key solution to the future energy crisis in AI computing, where the performance of optical switches plays a critical role. Conventional integrated optical switches typically suffer from poor robustness to voltage fluctuations, fabrication variations, and temperature drifts. These limitations necessitate complex high-prec

  72. Rozhin Yousefjani, Saif Al-Kuwari

    Stark-localized quantum probes have recently been shown to enable quantum-enhanced weak-field sensing with polynomial or super-polynomial scaling. In this paper, we show that the spatial geography of the encoded field can elevate this advantage to a genuine exponential scaling. We study a one-dimensional Stark probe subject to an exponential gradient profile

  73. Victor Chen, Ayden Coughlin, Michael D. Bond

    Automatically translating system software from C to Rust is an appealing but challenging problem, as it requires whole-program reasoning to satisfy Rust's ownership and borrowing discipline. A key enabling step in whole-program translation is interface translation, which produces Rust declarations for the C program's top-level declarations (i.e., structs and

  74. Arnav Goel, Pranjal A Chitale, Bhawna Paliwal, Bishal Santra

    User behavior in the real world is diverse, cross-domain, and spans long time horizons. Existing user modeling benchmarks however remain narrow, focusing mainly on short sessions and next-item prediction within a single domain. Such limitations hinder progress toward robust and generalizable user models. We present HORIZON, a new benchmark that reformulates

  75. Yifei Yan, Yankai Liao, Linqi Ye

    Deploying a humanoid robot to manipulate a new object has traditionally required one to two days of effort: data collection, manual annotation, 3D model acquisition, and model training. This paper presents an end-to-end rapid deployment pipeline that integrates three foundation-model components to shorten the onboarding cycle for a new object to approximatel

  76. Seungmin Lee, Jeonghwan Lee, Hyunkuk Lim, Sejoon Kim

    Recent text embedding models are often adapted to specialized domains via contrastive pre-finetuning (PFT) on a naive collection of scattered, heterogeneous tasks. However, this approach often introduces task-induced bias alongside domain knowledge, leading to uncontrolled representation shifts that distort the pretrained embedding geometry and cause substan

  77. Sheldon Paul, Izzat Alsmadi

    Assessing the security posture of modern computing systems typically requires the use of multiple specialized tools. These tools focus on different aspects such as configuration compliance, file integrity, and vulnerability exposure, and their outputs are often difficult to interpret collectively. This paper introduces the Unified Compliance Aggregator (UCA)

  78. Li Zheng, Xin Zhang, Shuyi He, Fei Li

    Accurate comprehension and controllable generation of emotion and rhetoric are pivotal for enhancing the reasoning capabilities of large language models (LLMs). Existing studies mostly rely on external optimizations, lacking in-depth exploration of internal representation mechanisms, thus failing to achieve fine-grained steering at the neuron level. A handfu

  79. Baichen Yu, Xuetong Li, Jing Zhou, Hansheng Wang

    Breast cancer is the most prevalent cancer in women worldwide. Histopathology image analysis serves as the gold standard for cancer diagnosis. In this regard, whole-slide imaging (WSI), a revolutionary technology in digital pathology, allows for ultrahigh-resolution tissue analysis. Despite its promise, WSI analysis faces significant computational challenges

  80. Hanlin Wang, Chak Tou Leong, Jian Wang, Wenjie Li

    Recent advancements in large language models (LLMs) have enabled agents to tackle complex embodied tasks through environmental interaction. However, these agents still make suboptimal decisions and perform ineffective actions, as they often overlook critical environmental feedback that differs from their internal beliefs. Through a formal probing analysis, w

  81. Boris Kriuk, Fedor Kriuk

    Standard risk models reduce the rich dependence structure of financial markets to scalar volatility estimates, discarding the topological information encoded in cross-asset correlation networks. We present ORCA (Online Regime Correlation Analyzer), an end-to-end framework that fuses spectral graph theory, random matrix theory, and supervised machine learning

  82. Pegah Golchian, Pauline Maier, Thomas Kocar, Marvin N. Wright

    Data scarcity challenges the development and implementation of innovative healthcare solutions. In geriatrics, fall-related injuries are a major cause of hospitalization, functional decline, and mortality in older adults. Optimizing post-operative discharge planning can mitigate these outcomes, but limited data hinders predictive model development. Here, we

  83. Sola Kim, Marco A. Janssen, Jieshu Wang, Ame Min-Venditti

    Federal agencies are increasingly deploying large language models (LLMs) to process public comments submitted during notice-and-comment rulemaking, the primary mechanism through which citizens influence federal regulation. Whether these systems treat all public input equally remains largely untested. Using a counterfactual design, we held comment content con

  84. Zhuoheng Li, Qingquan Lin, Checheng Yu, Qiangyu Chen

    High-DOF dexterous hands require compact actuation, rich sensing, and reliable thermal behavior, but conventional designs often occupy valuable in-hand space, increase end-effector mass, and suffer from heat accumulation near the hand. Remote tendon-driven actuation offers an alternative by relocating motors to the robot base or an external motor hub, thereb

  85. Priya Gurjar, Md Farhan Ishmam, Kenneth Marino

    Large language model (LLM) agents for sequential decision-making struggle to produce diverse outputs. This leads to insufficient exploration, suboptimal solutions, and repeated actions. Actions are generated at the sequence level, but existing sampling strategies, such as temperature scaling, introduce diversity at the token level, not at the sequence level.

  86. Rui Min, Liang Yao, Shiyu Miao, Shengxiang Xu

    A robust Multimodal Large Language Model (MLLM) for Earth Observation should maintain consistent interpretation and reasoning under realistic input variations. However, current Remote Sensing MLLMs fail to meet this requirement. Trained on carefully curated clean datasets, they learn brittle mappings that do not generalize to noisy conditions in operational

  87. Yi Xu, Yi-Zheng Fan

    The generalized Tur\'an number $\text{ex}(n, H, F)$ denotes the maximum number of copies of $H$ in an $n$-vertex $F$-free graph. Let $kK_{r+1}$ be the disjoint union of $k$ copies of the complete graph $K_{r+1}$. Recently, Gerbner determined $\text{ex}(n, K_{t},kK_{r+1})$ for all sufficiently large $n$. In this paper, we study a spectral analogue of this pro

  88. Kun Wang, Yiming Li, Mingcheng Qu, Aqiang Zhang

    Implicit spatial relations and deep semantic structures encoded in object attributes are crucial for procedural planning in embodied AI systems. However, existing approaches often over rely on the reasoning capabilities of vision language models (VLMs) themselves, while overlooking the rich structured semantic information that can be mined from multimodal in

  89. Vinil Pasupuleti, Shyalendar Reddy Allala, Siva Rama Krishna Varma Bayyavarapu, Shrey Tyagi

    Enterprise AI systems increasingly deploy multiple intelligent agents across mission-critical workflows that must satisfy hard policy constraints, bounded risk exposure, and comprehensive auditability (SOX, HIPAA, GDPR). Existing coordination methods - cooperative MARL, consensus protocols, and centralized planners - optimize expected reward while treating c

  90. Ziming Lin, Fang Han

    Double/debiased machine learning (DML) provides a general framework for inference with high-dimensional or otherwise complex nuisance parameters by combining Neyman-orthogonal scores with cross-fitting, thereby circumventing classical Donsker-type conditions in many modern machine-learning settings. Despite its strong empirical performance, bootstrap inferen

  91. Jiaqi Zhao, Fengwei Wang

    In the 47th IEEE Symposium on Security and Privacy (IEEE S&P 2026), Gao et al. proposed an efficient and user-friendly secure transformer inference framework, namely Euston. In Euston, a singular value decomposition-based matrix transmission protocol is designed to efficiently transmit input matrices, reducing communication bandwidth by approximately 2.8 tim

  92. Sunrit Chakraborty, XuanLong Nguyen

    In this paper, we develop a finite mixture of convolutional distributions, a statistical model to analyze continuous data distributed approximately on a mixture of low-dimensional affine subspaces. The observations are assumed independent and identically distributed from the mixture of distributions, where each component arises from a convolution of a distri

  93. Hui Gao, Xinming Wu, Jintao Li, Xiaoming Sun

    Seismic stratigraphic interpretation of shelf-edge clinothems is essential for revealing tectonic evolution, paleoclimate change, depositional dynamic conditions, and hydrocarbon generation and accumulation during basin filling. However, traditional interpretation methods remain labor-intensive, time-consuming, and highly subjective. Although AI-based method

  94. Shiyu He, Zhiman Chen, Yuqi Zhao, Neng Zhang

    The rapid expansion of the model context protocol (MCP) ecosystem enables large language model (LLM)-based agents to access a wide range of external tools via a standardized interface. However, identifying appropriate MCP servers for a specific development task remains challenging. Existing studies primarily focus on measuring the MCP ecosystem or optimizing

  95. Chun Wang, Chenfeng Wei, Chenyang Liu, Weihong Deng

    Personalized image aesthetics assessment (PIAA) aims to predict an individual user's subjective rating of an image, which requires modeling user-specific aesthetic preferences. Existing methods rely on historical user ratings for this modeling and therefore struggle when such data are unavailable. We address this zero-shot setting by using user profiles as c

  96. Badrinath Balasubramaniam, Vignesh Suresh, Benjamin Metcalf, Beiwen Li

    Unrecovered e-waste represents a significant economic loss. Hard disk drives (HDDs) comprise a valuable e-waste stream necessitating robotic disassembly. Automating the disassembly of HDDs requires holistic 3D sensing, scene understanding, and fastener localization, however current methods are fragmented, lack robust 3D sensing, and lack fastener localizatio

  97. Ramkishor Sharma, Samarth Majumdar, Divya Sachdeva

    Recently, a mechanism for generating astrophysically relevant magnetic fields via ultralight pseudoscalar dark matter, through the coupling term $g_{\phi \gamma} \phi F_{\mu \nu}\tilde{F}^{\mu\nu}$ in the Lagrangian density, was proposed in Brandenberger et al (2026) (see Ref. 1). In this scenario, the electromagnetic fields are amplified through the phenome

  98. Alexandre Linhares

    Project Yanasse presents a method for discovering new proofs of theorems in one area of mathematics by transferring proof strategy patterns (e.g., Lean 4 tactic invocation patterns) from a structurally distant area. The system extracts tactic usage distributions across 27 top-level areas of Mathlib (217,133 proof states), computes z-scores to identify tactic

  99. Kejia Bian, Meixia Tao, Jianhua Mo, Zhiyong Chen

    The success of large foundation models is catalyzing a new paradigm for AI-native 6G network design: wireless foundation models for physical-layer design. However, existing models often operate on channel state information (CSI) in the spatial-temporal-frequency (STF) domain, where multipath components are superimposed and structurally entangled. This hinder

  100. Qingwei Lin

    Conditional depth execution routes a subset of tokens through a lightweight cheap FFN while the remainder execute the standard full FFN at each controlled layer. The central difficulty is gate training: the gate decision must propagate through many layers before it influences the language modeling (LM) loss, so the resulting gradients are weak and noisy. Aux