Skip to content

May 2024 arXiv papers — page 46

Showing 4,5014,600 of 20,894 papers

  1. Hastings Greer, Lin Tian, Francois-Xavier Vialard, Roland Kwitt

    Image registration estimates spatial correspondences between a pair of images. These estimates are typically obtained via numerical optimization or regression by a deep network. A desirable property of such estimators is that a correspondence estimate (e.g., the true oracle correspondence) for an image pair is maintained under deformations of the input image

  2. Daniel A. Turolla Vanzella

    Although we lack complete understanding of quantum aspects of gravitation, it is usually agreed, using general arguments, that a final quantum gravity theory will endow space and time with some (fundamental or effective) notion of discreteness. This granular character is supposed to lie on space and time scales of $l_P \sim 10^{-33}$ cm and $\tau_P\sim 10^{-

  3. Ye He, Alireza Mousavi-Hosseini, Krishnakumar Balasubramanian, Murat A. Erdogdu

    We study the complexity of heavy-tailed sampling and present a separation result in terms of obtaining high-accuracy versus low-accuracy guarantees i.e., samplers that require only $O(\log(1/\varepsilon))$ versus $\Omega(\text{poly}(1/\varepsilon))$ iterations to output a sample which is $\varepsilon$-close to the target in $\chi^2$-divergence. Our results a

  4. Xiaomin Lv, Chong Lai, Liya Ding, Maode Lai

    Large models have become mainstream, yet their applications in digital pathology still require exploration. Meanwhile renal pathology images play an important role in the diagnosis of renal diseases. We conducted image segmentation and paired corresponding text descriptions based on 60 books for renal pathology, clustering analysis for all image and text des

  5. Kai Jia, Martin Rinard

    We study rational agents with different perception capabilities in strategic games. We focus on a class of one-shot limited-perception games. These games extend simultaneous-move normal-form games by presenting each player with an individualized perception of all players' payoff functions. The accuracy of a player's perception is determined by the player's c

  6. Xunpeng Huang, Difan Zou, Yi-An Ma, Hanze Dong

    Stochastic gradients have been widely integrated into Langevin-based methods to improve their scalability and efficiency in solving large-scale sampling problems. However, the proximal sampler, which exhibits much faster convergence than Langevin-based algorithms in the deterministic setting Lee et al. (2021), has yet to be explored in its stochastic variant

  7. S. B. Samuel, Z. Gedik

    Symmetric Informationally Complete Positive Operator-Valued Measures (SIC-POVMs) have been constructed in many dimensions using the Weyl-Heisenberg group. In the quantum information community, it is commonly believed that SCI-POVMs exist in all dimensions; however, the general proof of their existence is still an open problem. The Bloch sphere representation

  8. Dongyan Huo, Yixuan Zhang, Yudong Chen, Qiaomin Xie

    In this work, we investigate stochastic approximation (SA) with Markovian data and nonlinear updates under constant stepsize $\alpha>0$. Existing work has primarily focused on either i.i.d. data or linear update rules. We take a new perspective and carefully examine the simultaneous presence of Markovian dependency of data and nonlinear update rules, delinea

  9. Jeonghwan Cheon, Sang Wan Lee, Se-Bum Paik

    The brain prepares for learning even before interacting with the environment, by refining and optimizing its structures through spontaneous neural activity that resembles random noise. However, the mechanism of such a process has yet to be thoroughly understood, and it is unclear whether this process can benefit the algorithm of machine learning. Here, we st

  10. Peiyu Yu, Dinghuai Zhang, Hengzhi He, Xiaojian Ma

    Noise Contrastive Estimation (NCE) has fueled major breakthroughs in representation learning and generative modeling. Yet a long-standing challenge remains: accurately estimating ratios between distributions that differ substantially, which significantly limits the applicability of NCE on modern high-dimensional and multimodal datasets. We revisit this probl

  11. Md Zobaer Islam, Ethan Abele, Fahim Ferdous Hossain, Arsalan Ahmad

    Channel turbulence is a formidable obstacle for free-space optical (FSO) communication. Anticipation of turbulence levels is highly important for mitigating disruptions but has not been demonstrated without dedicated, auxiliary hardware. We show that machine learning (ML) can be applied to raw FSO data streams to rapidly predict channel turbulence levels wit

  12. Abrar Alali, Stephan Olariu

    Accommodating pedestrians crossing midblock has been shown to have harmful environmental consequences because of increased fuel consumption and CO2 emissions. Somewhat surprisingly, no studies were devoted to mitigating the environmental impact of midblock crossing. Our main contribution is to propose schemes that mitigate the increased fuel consumption and

  13. H. Jung, S. Dong, D. Zahn, T. Vasileiadis

    We study monolayer WSe2 using ultrafast electron diffraction. We introduce an approach to quantitatively extract atomic-site-specific information, providing an element-specific view of incoherent atomic vibrations following femtosecond excitation. Via differences between W and Se vibrations, we identify stages in the nonthermal evolution of the lattice. Comb

  14. Jianxun Hu, Huazhong Ke, Changzheng Li, Zhitong Su

    We estimate an upper bound of the spectral radius of a linear operator on the quantum cohomology of the toric Fano manifolds $\mathbb{P}_{\mathbb{P}^{n}}(\mathcal{O}\oplus\mathcal{O}(3))$. This provides a negative answer to Galkin's lower bound conjecture.

  15. Anirban Kundu, Won Kyung Seong, S. Kamal Jalali, Nicola M. Pugno

    We report the scientific and technical queries regarding the article reported by Kim et al.1 on the mechanical properties of graphene-poly(methyl methacrylate) (PMMA) composites. Our analysis finds that the current experimental data is insufficient to fully support the conclusions presented in the article. We suggest the enhancement in Youngs modulus and str

  16. Lijun Yu

    Advancements in language foundation models have primarily fueled the recent surge in artificial intelligence. In contrast, generative learning of non-textual modalities, especially videos, significantly trails behind language modeling. This thesis chronicles our endeavor to build multi-task models for generating videos and other modalities under diverse cond

  17. Awni Altabaa, John Lafferty

    Relational reasoning is a central component of generally intelligent systems, enabling robust and data-efficient inductive generalization. Recent empirical evidence shows that many existing neural architectures, including Transformers, struggle with tasks requiring relational reasoning. In this work, we distinguish between two types of information: sensory i

  18. Fanchen Bu, Ruochen Yang, Paul Bogdan, Kijung Shin

    Desirable random graph models (RGMs) should (i) reproduce common patterns in real-world graphs (e.g., power-law degrees, small diameters, and high clustering), (ii) generate variable (i.e., not overly similar) graphs, and (iii) remain tractable to compute and control graph statistics. A common class of RGMs (e.g., Erdos-Renyi and stochastic Kronecker) output

  19. G. Grover, N. D. R. Bhat, S. McSweeney, C. P. Lee

    The phenomenon of pulsar nulling, where pulsars temporarily and stochastically cease their radio emission, is thought to be indicative of a `dying' pulsar, where radio emission ceases entirely. Here we report the discovery of a long-period pulsar, PSR J0452-3418, from the ongoing Southern-sky MWA Rapid Two-meter (SMART) pulsar survey. The pulsar has a rotati

  20. Johannes C. Bayer, Fredrik Brange, Adrian Schmidt, Timo Wagner

    We experimentally demonstrate the real-time detection and control of correlated charge tunneling in a dynamically driven quantum dot. Specifically, we measure the joint distribution of waiting times between tunneling charges and show that the waiting times for holes may be strongly correlated due to the periodic drive and the Coulomb interactions on the dot,

  21. Akihiro Goto

    Lehmer conjectured that Ramanujan's tau function never vanishes. As a variation of this conjecture, it is proved that \begin{equation*} \tau(n)\neq \pm \ell, \pm 2\ell, \pm 2\ell^2, \end{equation*} where $\ell<100$ is an odd prime, by Balakrishnan, Ono, Craig, Tsai and many people. We have proved that \begin{equation*} \tau(n)\neq \pm \ell, \pm 2\ell, \pm 4\

  22. Ayano Nakamura, Shinichi Nishihaya, Hiroaki Ishizuka, Markus Kriener

    For over a century, the Hall effect, a transverse effect under out-of-plane magnetic field or magnetization, has been a cornerstone for magnetotransport studies and applications. Modern theoretical formulation based on the Berry curvature has revealed the potential that even in-plane magnetic field can induce anomalous Hall effect, but its experimental demon

  23. Kazuhiro Ishige, Qing Liu, Paolo Salani

    In this paper, we provide a new PDE proof for the celebrated Borell--Brascamp--Lieb inequality. Our approach reveals a deep connection between the Borell--Brascamp--Lieb inequality and properties of diffusion equations of porous medium type pertaining to the large time asymptotics and preservation of a generalized concavity of the solutions. We also recover

  24. Yu Wang, Ruihan Wu, Zexue He, Xiusi Chen

    Large language models show impressive abilities in memorizing world knowledge, which leads to concerns regarding memorization of private information, toxic or sensitive knowledge, and copyrighted content. We introduce the problem of Large Scale Knowledge Washing, focusing on unlearning an extensive amount of factual knowledge. Previous unlearning methods usu

  25. Pierre Tholoniat, Kelly Kostopoulou, Peter McNeely, Prabhpreet Singh Sodhi

    With the impending removal of third-party cookies from major browsers and the introduction of new privacy-preserving advertising APIs, the research community has a timely opportunity to assist industry in qualitatively improving the Web's privacy. This paper discusses our efforts, within a W3C community group, to enhance existing privacy-preserving advertisi

  26. Yashas Annadani, Panagiotis Tigas, Stefan Bauer, Adam Foster

    We present Causal Amortized Active Structure Learning (CAASL), an active intervention design policy that can select interventions that are adaptive, real-time and that does not require access to the likelihood. This policy, an amortized network based on the transformer, is trained with reinforcement learning on a simulator of the design environment, and a re

  27. Sabry G Moustafa

    Convergence of path integral simulations requires a substantial number of beads when quantum effects are significant. Traditional Trotter scaling approaches estimate the continuum limit through extrapolation, however they are restricted to the asymptotic behavior near this limit. We introduce an efficient extrapolation approach for thermodynamic properties o

  28. Chinmay Maheshwari, Kshitij Kulkarni, Manxi Wu, Shankar Sastry

    We propose an adaptive incentive mechanism that learns the optimal incentives in environments where players continuously update their strategies. Our mechanism updates incentives based on each player's externality, defined as the difference between the player's marginal cost and the operator's marginal cost at each time step. The proposed mechanism updates t

  29. Chong Chen, Yingmin Liu, Yu Ding, Matthew Tong

    Background: Accelerated real-time cine (RT-Cine) imaging enables cardiac function assessment without the need for breath-holding. However, when performed during in-magnet exercise, RT-Cine images may exhibit significant motion artifacts. Methods: By projecting the time-averaged images to the subspace spanned by the coil sensitivity maps, we propose a coil re

  30. Vinamra Benara, Chandan Singh, John X. Morris, Richard Antonello

    Large language models (LLMs) have rapidly improved text embeddings for a growing array of natural-language processing tasks. However, their opaqueness and proliferation into scientific domains such as neuroscience have created a growing need for interpretability. Here, we ask whether we can obtain interpretable embeddings through LLM prompting. We introduce

  31. Bertrand Marchand, Nadia Tahiri, Olivier Tremblay-Savard, Manuel Lafond

    In this paper, we lay the groundwork on the comparison of phylogenetic networks based on edge contractions and expansions as edit operations, as originally proposed by Robinson and Foulds to compare trees. We prove that these operations connect the space of all phylogenetic networks on the same set of leaves, even if we forbid contractions that create cycles

  32. Paolo Glorioso, Quentin Anthony, Yury Tokpanov, James Whittington

    In this technical report, we present Zamba, a novel 7B SSM-transformer hybrid model which achieves competitive performance against leading open-weight models at a comparable scale. Zamba is trained on 1T tokens from openly available datasets and is the best non-transformer model at this scale. Zamba pioneers a unique architecture combining a Mamba backbone w

  33. Christine P Lee, Min Kyung Lee, Bilge Mutlu

    Increasing evidence suggests that many deployed AI systems do not sufficiently support end-user interaction and information needs. Engaging end-users in the design of these systems can reveal user needs and expectations, yet effective ways of engaging end-users in the AI explanation design remain under-explored. To address this gap, we developed a design met

  34. Christine P Lee, Pragathi Praveena, Bilge Mutlu

    Robots in real-world environments continuously engage with multiple users and encounter changes that lead to unexpected conflicts in fulfilling user requests. Recent technical advancements (e.g., large-language models (LLMs), program synthesis) offer various methods for automatically generating repair plans that address such conflicts. In this work, we under

  35. Inhee Lee, Jiefu Cen, Oleksandr Molchanov, Shi Feng

    CrX$_3$ (X = Cl, Br, I) have the same crystal structure and Hamiltonian but different ligand spin-orbit coupling (SOC) constant $\lambda_X$, providing excellent material platform exploring for exotic two-dimensional (2D) spin orders. Their microscopic mechanism underlying 2D spin physics remain unestablished, along with experimental corroboration of Kitaev e

  36. Xueqing Zhang, Junkai Zhang, Ka-Ho Chow, Juntao Chen

    This demo paper examines the susceptibility of Federated Learning (FL) systems to targeted data poisoning attacks, presenting a novel system for visualizing and mitigating such threats. We simulate targeted data poisoning attacks via label flipping and analyze the impact on model performance, employing a five-component system that includes Simulation and Dat

  37. Karan Goyal, Mayank Goel, Vikram Goyal, Mukesh Mohania

    Citing pertinent literature is pivotal to writing and reviewing a scientific document. Existing techniques mainly focus on the local context or the global context for recommending citations but fail to consider the actual human citation behaviour. We propose SymTax, a three-stage recommendation architecture that considers both the local and the global contex

  38. Carmen Gómez-Fayrén, Tomás Ortín, Matteo Zatti

    We show that, in presence of isometries and non-trivial topology, the Einstein--Hilbert action is invariant under certain transformations of the metric which are not diffeomorphisms. These transformations are similar to the higher-form symmetries of field theories with $p$-form fields. In the context of toroidal Kaluza--Klein compactifications, we show that

  39. Tianlong Wang, Xianfeng Jiao, Yinghao Zhu, Zhongzhi Chen

    Recent studies have indicated that Large Language Models (LLMs) harbor an inherent understanding of truthfulness, yet often fail to consistently express it and generate false statements. This gap between "knowing" and "telling" poses a challenge for ensuring the truthfulness of generated content. Inspired by recent work on the practice of encoding human-inte

  40. Pier Domenico Lamberti, Vitaly Moroz

    We study sub and supersolutions for the $p$-Laplace type elliptic equation of the form $$-\Delta_p u-V|u|^{p-2}u=0\quad\text{in $\Omega$},$$ where $\Omega$ is a radially symmetric domain in ${\mathbb{R}}^N$ and $V(x)\ge 0$ is a continuous potential such that the solutions of the equation satisfy the comparison principle on bounded subdomains of $\Omega$. In

  41. Tom Benhamou, Alejandro Poveda

    We develop the non-normal variations of two classical Prikry-type forcings; namely, Magidor and Radin forcings. We generalize the fact that the non-normal Prikry forcing is a projection of the extender-based to a coordinate of the extender to our forcing and the Radin/Magidor-Radin-extender-based forcing from \cite{CarmiMagidorRadin,CarmiRadin}. Then, we sho

  42. Natalia Amburg, Ilya Tolstukhin

    We study the three-point quantum $\mathfrak{sl_2}$-Gaudin model. In this case the compactification of the parameter space is $\overline{M_{0,4}(\mathbb{C})}$, which is the Riemann sphere. We analyze sphere coverings by the joint spectrum of the Gaudin Hamiltonians treating them as algebraic curves. We write equations of these curves as determinants of tridia

  43. Peiran Yao, Denilson Barbosa

    Open-domain question answering (Open-QA) is a common task for evaluating large language models (LLMs). However, current Open-QA evaluations are criticized for the ambiguity in questions and the lack of semantic understanding in evaluators. Complex evaluators, powered by foundation models or LLMs and pertaining to semantic equivalence, still deviate from huma

  44. Tong Shi, Xuri Ge, Joemon M. Jose, Nicolas Pugeault

    Capturing complex temporal relationships between video and audio modalities is vital for Audio-Visual Emotion Recognition (AVER). However, existing methods lack attention to local details, such as facial state changes between video frames, which can reduce the discriminability of features and thus lower recognition accuracy. In this paper, we propose a Detai

  45. Mustafa Shukor, Matthieu Cord

    Large Language Models (LLMs) have demonstrated impressive performance on multimodal tasks, without any multimodal finetuning. They are the building block for Large Multimodal Models, yet, we still lack a proper understanding of their success. In this work, we expose frozen LLMs to image, video, audio and text inputs and analyse their internal representation

  46. Ruben Lier

    Measuring lift force on symmetrically shaped obstacles immersed in laminar flow is the quintessential way of signalling odd viscosity. For flow past cylinders, such a lift force does not arise when incompressibility and no-slip boundary conditions hold, whereas for spheres, a lift force was found in Stokes flow, applying to cases where the Reynolds number is

  47. Tianji Cai, Junyi Cheng, Nathaniel Craig, Giacomo Koszegi

    How can one fully harness the power of physics encoded in relativistic $N$-body phase space? Topologically, phase space is isomorphic to the product space of a simplex and a hypersphere and can be equipped with explicit coordinates and a Riemannian metric. This natural structure that scaffolds the space on which all collider physics events live opens up new

  48. Azim Akhtarshenas, Navid Ayoobi, David Lopez-Perez, Ramin Toosi

    Optimizing the design, performance, and resource efficiency of wireless networks (WNs) necessitates the ability to discern Line of Sight (LoS) and Non-Line of Sight (NLoS) scenarios across diverse applications and environments. Unmanned Aerial Vehicles (UAVs) exhibit significant potential in this regard due to their rapid mobility, aerial capabilities, and p

  49. Pegah Golestaneh, Mahsa Taheri, Johannes Lederer

    Neural networks have become standard tools in many areas, yet many important statistical questions remain open. This paper studies the question of how much data are needed to train a ReLU feed-forward neural network. Our theoretical and empirical results suggest that the generalization error of ReLU feed-forward neural networks scales at the rate $1/\sqrt{n}

  50. Elliot M. Miller, Tat Chung D. Chan, Carlos Montes-Matamoros, Omar Sharif

    Many neurodegenerative diseases (NDs) are characterized by the slow spatial spread of toxic protein species in the brain. The toxic proteins can induce neuronal stress, triggering the Unfolded Protein Response (UPR), which slows or stops protein translation and can indirectly reduce the toxic load. However, the UPR may also trigger processes leading to apopt

  51. Chongjun Ouyang, Yuanwei Liu, Xingqi Zhang

    The concept of aperture selection is proposed for continuous aperture array (CAPA)-based communications. The achieved performance is analyzed in an uplink scenario by considering both line-of-sight (LoS) and non-line-of-sight (NLoS) scenarios. In the LoS scenario, the optimal selection strategy is demonstrated to follow the nearest neighbor criterion, and th

  52. Michał Strada, Sebastian Ernst, Jacek Szybowski, Konrad Kułakowski

    Most decision-making models, including the pairwise comparison method, assume the decision-makers honesty. However, it is easy to imagine a situation where a decision-maker tries to manipulate the ranking results. This paper presents three simple manipulation methods in the pairwise comparison method. We then try to detect these methods using appropriately c

  53. Manish Saini, Melvin Paul Jacob, Minh Nguyen, Nico Hochgeschwender

    When performing manipulation-based activities such as picking objects, a mobile robot needs to position its base at a location that supports successful execution. To address this problem, prominent approaches typically rely on costly grasp planners to provide grasp poses for a target object, which are then are then analysed to identify the best robot placeme

  54. Maruša Lekše

    A walk of length $n$ in a graph is consistent if there exists an automorphism of the graph that maps the initial $n-1$ vertices to the final $n-1$ vertices of the walk. In this paper we find some sufficient conditions for a consistent walk in an arc-transitive graph to have a trivial pointwise stabilizer. We show that in that case, the size of the smallest g

  55. Chongjun Ouyang, Yuanwei Liu, Xingqi Zhang

    The performance of continuous aperture array (CAPA)-based wireless communications is analyzed in an uplink scenario. An analytical framework is proposed to characterize uplink CAPA-based transmission using electromagnetic field theories. On this basis, new expressions are derived for the channel capacity in a single-user scenario and the sum-rate capacity in

  56. Daniel López-Fernández, Ricardo Vergaz

    Contribution: The combination of ChatGPT with traditional learning resources is very effective in computer science education. High-performing students are the ones who are using ChatGPT the most. So, a new digital trench could be rising between these students and those with lower degree of fundamentals and worse prompting skills, who may not take advantage o

  57. Tuan Q. Do

    We point out that the area of event horizon of Kleinian black hole is infinite due to the fact that its event horizon is not a sphere but a hyperboloid. Therefore, the usual interpretations of Schwarzschild black hole might not be applicable to the Kleinian black hole.

  58. Rem Sadykhov, Geoffrey Goodell, Philip Treleaven

    This paper presents a theoretical extension of the DeTEcT framework proposed by Sadykhov et al., DeTEcT, where a formal analysis framework was introduced for modelling wealth distribution in token economies. DeTEcT is a framework for analysing economic activity, simulating macroeconomic scenarios, and algorithmically setting policies in token economies. This

  59. Xavier Riley, Simon Dixon

    The Charlie Parker Omnibook is a cornerstone of jazz music education, described by pianist Ethan Iverson as "the most important jazz education text ever published". In this work we propose a new transcription pipeline and explore the extent to which state of the art music technology is able to reconstruct these scores directly from the audio without human in

  60. C. Granier, S. S. Cerri, F. Jenko

    We perform 3D3V hybrid-Vlasov simulations of turbulence with quasi-isotropic, compressible injection near ion scales to mimic the Earth's magnetosheath plasma, and investigate the novel electron-only reconnection, recently observed by the NASA's MMS mission, and its impact on ion heating. Retaining electron inertia in the generalized Ohm s law enables collis

  61. Christian Makaya, Keith Grueneberg, Bongjun Ko, David Wood

    Computing at the edge is increasingly important as Internet of Things (IoT) devices at the edge generate massive amounts of data and pose challenges in transporting all that data to the Cloud where they can be analyzed. On the other hand, harnessing the edge data is essential for offering cognitive applications, if the challenges, such as device capabilities

  62. Rohan Pandey

    Past work has established scaling laws that predict the performance of a neural language model (LM) as a function of its parameter count and the number of tokens it's trained on, enabling optimal allocation of a fixed compute budget. Are these scaling laws agnostic to training data as some prior work suggests? We generate training datasets of varying complex

  63. Abid Faisal Ayon, S M Maksudul Alam

    Facial Recognition is a technique, based on machine learning technology that can recognize a human being analyzing his facial profile, and is applied in solving various types of realworld problems nowadays. In this paper, a common real-world problem, finding a missing person has been solved in a secure and effective way with the help of facial recognition te

  64. Mustapha Raissouli, Mohamed Chergui

    In this article, we define a special function called the Bigamma function. It provides a generalization of Euler's gamma function. Several algebraic properties of this new function are studied. In particular, results linking this new function to the standard Beta function have been provided. We have also established inequalities, which allow to approximate t

  65. Ashkan Vedadi Gargary, Emiliano De Cristofaro

    Federated Learning (FL) has emerged as a solution for distributed systems that allow clients to train models on their data and only share models instead of local data. Generative Models are designed to learn the distribution of a dataset and generate new data samples that are similar to the original data. Many prior works have tried proposing Federated Gener

  66. Amir Saeidi, Shivanshu Verma, Aswin RRV, Kashif Rasul

    Reinforcement Learning with Human Feedback (RLHF) enhances the alignment of Large Language Models (LLMs). However, its limitations have led to the development of Direct Preference Optimization (DPO), an RL-free approach designed to overcome these shortcomings. While studies have shown that DPO improves instruction-following capabilities, it negatively impact

  67. Taewan Kim, Abhinav G. Kamath, Niyousha Rahimi, Jasper Corleis

    This paper presents a numerical optimization algorithm for generating approach and landing trajectories for a six-degree-of-freedom (6-DoF) aircraft. We improve on the existing research on aircraft landing trajectory generation by formulating the trajectory optimization problem with additional real-world operational constraints, including 6-DoF aircraft dyna

  68. Rafael Bailo, José A. Carrillo, David Gómez-Castro

    This is a survey article based on the content of the plenary lecture given by Jos\'e A. Carrillo at the ICIAM23 conference in Tokyo. It is devoted to produce a snapshot of the state of the art in the analysis, numerical analysis, simulation, and applications of the vast area of aggregation-diffusion equations. We also discuss the implications in mathematical

  69. Alex C. Dantas, Junio R. Oliveira, Tulio M. G. Santos

    A finitely generated group is said to be an automata group if it admits a faithful self-similar finite-state representation on some regular $m$-tree. We prove that if $G$ is a subgroup of an automata group, then for each finitely generated abelian group $A$, the wreath product $A \wr G$ is a subgroup of an automata group. We obtain, for example, that $C_2 \w

  70. Yuanchao Li, Pinzhen Chen, Peter Bell, Catherine Lai

    ASR remains unsatisfactory in scenarios where the speaking style diverges from that used to train ASR systems, resulting in erroneous transcripts. To address this, ASR Error Correction (AEC), a post-ASR processing approach, is required. In this work, we tackle an understudied issue: the Low-Resource Out-of-Domain (LROOD) problem, by investigating crossmodal

  71. Alvaro García, Esteban Cañibano

    The digitization of the productive ecosystem related to human-machine interaction has become a priority for small and medium-sized enterprises. Particularly to face the challenges of Industry 4.0 and advanced digital skills in the workplace. From the research point of view, digitization opens a global scenario for the generation of opportunities for learning

  72. Eugenio Tufino, Stefano Oss, Micol Alemani

    In this paper, we detail the integration of Python data analysis into a first-year physics laboratory course, a task accomplished without significant alterations to the existing course structure. We introduced tailored laboratory computational learning goals and designed activities to address them. We emphasise the development and application of Jupyter Note

  73. Nikola Zubić, Federico Soldá, Aurelio Sulser, Davide Scaramuzza

    Despite their successes, deep learning models struggle with tasks requiring complex reasoning and function composition. We present a theoretical and empirical investigation into the limitations of Structured State Space Models (SSMs) and Transformers in such tasks. We prove that one-layer SSMs cannot efficiently perform function composition over large domain

  74. Nikodem Popławski

    The McVittie metric does not describe a physical black hole in an expanding Universe because the curvature scalar and pressure at its event horizon are infinite. We show that extending this metric to an inhomogeneous scale factor, which depends on both the time and radial coordinate, removes those infinities by imposing at the horizon the constancy of the Hu

  75. Jiachen Chen, Danyang Huang, Liyuan Wang, Kathryn L. Lunetta

    Node classification is a fundamental task, but obtaining node classification labels can be challenging and expensive in many real-world scenarios. Transfer learning has emerged as a promising solution to address this challenge by leveraging knowledge from source domains to enhance learning in a target domain. Existing transfer learning methods for node class

  76. Zhan Su, Fengran Mo, Prayag Tiwari, Benyou Wang

    In multi-task learning, the conventional approach involves training a model on multiple tasks simultaneously. However, the training signals from different tasks can interfere with one another, potentially leading to \textit{negative transfer}. To mitigate this, we investigate if modular language models can facilitate positive transfer and systematic generali

  77. W. S. Ożański, W. Zajączkowski

    We consider the axisymmetric Navier-Stokes equations in a finite cylinder $\Omega\subset\mathbb{R}^3$. We assume that $v_r$, $v_\varphi$, $\omega_\varphi$ vanish on the lateral boundary $\partial \Omega$ of the cylinder, and that $v_z$, $\omega_\varphi$, $\partial_z v_\varphi$ vanish on the top and bottom parts of the boundary $\partial \Omega$, where we use

  78. Hellina Hailu Nigatu, John Canny, Sarah E. Chasins

    Online Knowledge Repositories (OKRs) like Wikipedia offer communities a way to share and preserve information about themselves and their ways of living. However, for communities with low-resourced languages -- including most African communities -- the quality and volume of content available are often inadequate. One reason for this lack of adequate content c

  79. Christos Sourdis

    We consider strongly coupled competitive elliptic systems of Gross-Pitaevskii type that arise in the study of two-component Bose-Einstein condensates, in general smooth bounded domains of $\mathbb{R}^N$, $N\geq 1$. As the coupling parameter tends to infinity, solutions that remain uniformly bounded are known to converge to a segregated limiting profile, with

  80. Dirk Tasche

    The purpose of class distribution estimation (also known as quantification) is to determine the values of the prior class probabilities in a test dataset without class label observations. A variety of methods to achieve this have been proposed in the literature, most of them based on the assumption that the distributions of the training and test data are rel

  81. Xijie Huang, Xinyuan Wang, Hantao Zhang, Yinghao Zhu

    Security concerns related to Large Language Models (LLMs) have been extensively explored, yet the safety implications for Multimodal Large Language Models (MLLMs), particularly in medical contexts (MedMLLMs), remain insufficiently studied. This paper delves into the underexplored security vulnerabilities of MedMLLMs, especially when deployed in clinical envi

  82. Eugenio Giannelli, Noelia Rizo, A. A. Schaeffer Fry, Carolina Vallejo

    We characterize when a finite group G possesses a Sylow 3-subgroup P with abelianization of order 9 in terms of the number of height zero characters lying in the principal 3-block of G, settling a conjecture put forward by Navarro, Sambale, and Tiep in 2018. Along the way, we show that a recent result by Laradji on the number of character of height zero in a

  83. Chao Li, Jinwei Zhang, Hang Zhang, Jiahao Li

    Purpose: To develop a pipeline for motion artifact correction in mGRE and quantitative susceptibility mapping (QSM). Methods: Deep learning is integrated with autofocus to improve motion artifact suppression, which is applied QSM of patients with Parkinson's disease (PD). The estimation of affine motion parameters in the autofocus method depends on signal-to

  84. Hongjie Chen, Jingqiu Ding, Yiding Hua, David Steurer

    We give the first polynomial-time, differentially node-private, and robust algorithm for estimating the edge density of Erd\H{o}s-R\'enyi random graphs and their generalization, inhomogeneous random graphs. We further prove information-theoretical lower bounds, showing that the error rate of our algorithm is optimal up to logarithmic factors. Previous algori

  85. Stepan L. Kuznetsov, Alexander Okhotin

    A new family of categorial grammars is proposed, defined by enriching basic categorial grammars with a conjunction operation. It is proved that the formalism obtained in this way has the same expressive power as conjunctive grammars, that is, context-free grammars enhanced with conjunction. It is also shown that categorial grammars with conjunction can be na

  86. Piyush Jha, Prithwish Jana, Pranavkrishna Suresh, Arnav Arora

    Large Language Models (LLMs) have transformed AI but often struggle with tasks that require domain-specific reasoning and logical alignment. Traditional fine-tuning methods do not leverage the vast amount of symbolic domain-knowledge available to us via symbolic reasoning tools (e.g., provers), and are further limited by sparse rewards and unreliable reward

  87. Simon Segert

    Consider the following probability puzzle: A fair coin is flipped n times. For each HT in the resulting sequence, Bob gets a point, and for each HH Alice gets a point. Who is more likely to win? We provide a proof that Bob wins more often for every n>=3. As a byproduct, we derive the asymptotic form of the difference in win probabilities, and obtain an effic

  88. Thomas Manteaux, David Rodríguez-Martínez, Raj Thilak Rajan

    Efficient path planning is key for safe autonomous navigation over complex and unknown terrains. Lunar Zebro (LZ), a project of the Delft University of Technology, aims to deploy a compact rover, no larger than an A4 sheet of paper and weighing not more than 3 kilograms. In this work, we introduce a Robust Artificial Potential Field (RAPF) algorithm, a new p

  89. Yeachan Park, Minseok Kim, Yeoneung Kim

    We propose novel methodologies aimed at accelerating the grokking phenomenon, which refers to the rapid increment of test accuracy after a long period of overfitting as reported in~\cite{power2022grokking}. Focusing on the grokking phenomenon that arises in learning arithmetic binary operations via the transformer model, we begin with a discussion on data au

  90. Jiaxi Yu, Ashley J. Ross, Antoine Rocher, Otávio Alves

    Dark Energy Spectroscopic Instrument (DESI) uses more than 2.4 million Emission Line Galaxies (ELGs) for 3D large-scale structure (LSS) analyses in its Data Release 1 (DR1). Such large statistics enable thorough research on systematic uncertainties. In this study, we focus on spectroscopic systematics of ELGs. The redshift success rate ($f_{\rm goodz}$) is t

  91. Hellina Hailu Nigatu, Inioluwa Deborah Raji

    Online social media platforms such as YouTube have a wide, global reach. However, little is known about the experience of low-resourced language speakers on such platforms; especially in how they experience and navigate harmful content. To better understand this, we (1) conducted semi-structured interviews (n=15) and (2) analyzed search results (n=9313), rec

  92. Keun Soo Yim

    This paper presents a framework that selectively triggers security reviews for incoming source code changes. Functioning as a review bot within a code review service, the framework can automatically request additional security reviews at pre-submit time before the code changes are submitted to a source code repository. Because performing such secure code rev

  93. Inha Cha, Ajit G. Pillai, Richmond Y. Wong

    This paper introduces Ethics Pathways, a design activity aimed at understanding HCI and design researchers' ethics engagements and flows during their research process. Despite a strong ethical commitment in these fields, challenges persist in grasping the complexity of researchers' engagement with ethics -- practices conducted to operationalize ethics -- in

  94. Andrew Lane, Natasha Morrison

    Given graphs $G, H$ and an integer $q \ge 2$, the generalized Ramsey number, denoted $r(G,H,q)$, is the minimum number of colours needed to edge-colour $G$ such that every copy of $H$ receives at least $q$ colours. In this paper, we prove that for a fixed integer $k \ge 3$, we have $r(K_n,C_k,3) = n/(k-2)+o(n)$. This generalises work of Joos and Muybayi, who

  95. Alexandre Girard, H. Harry Asada

    This paper present a novel dual-speed actuator adapted to robotics. In many applications, robots have to bear large loads while moving slowly and also have to move quickly through the air with almost no load. This lead to conflicting requirements for their actuators. Multiple gear ratios address this issue by allowing an effective use of power over a wide ra

  96. Amit Surana, Abeynaya Gnanasekaran

    We present a novel variational quantum framework for linear partial differential equation (PDE) constrained optimization problems. Such problems arise in many scientific and engineering domains. For instance, in aerodynamics, the PDE constraints are the conservation laws such as momentum, mass and energy balance, the design variables are vehicle shape parame

  97. Linglong Qian, Yiyuan Yang, Wenjie Du, Jun Wang

    This study investigates the impact of masking strategies on time series imputation models in healthcare settings. While current approaches predominantly rely on random masking for model evaluation, this practice fails to capture the structured nature of missing patterns in clinical data. Using the PhysioNet Challenge 2012 dataset, we analyse how different ma

  98. David Benisty, David Vasak, Jürgen Struckmeier, Horst Stöcker

    Dark energy (and its simplest model, the Cosmological Constant or $\Lambda$) acts as a repulsive force that opposes gravitational attraction. Assuming galaxies maintain a steady state over extended periods, the estimated upper limit on $\Lambda$ studies its pushback to the attractive gravitational force of dark matter. From the SPARC dataset, we select galax

  99. Wenjian Hao, Devesh Upadhyay, Shaoshuai Mou

    This paper proposes a data-driven framework to learn a finite-dimensional approximation of a Koopman operator for approximating the state evolution of a dynamical system under noisy observations. To this end, our proposed solution has two main advantages. First, the proposed method only requires the measurement noise to be bounded. Second, the proposed metho

  100. Jakob Glas

    By developing a suitable version of the circle method, we show that the space of degree $e$ rational curves on a smooth hypersurface of degree $d$ has only canonical singularities provided its dimension is sufficiently large with respect to $e$ and $d$.