Skip to content

October 2020 arXiv papers — page 40

Showing 3,9014,000 of 16,697 papers

  1. Shafin Rahman, Sejuti Rahman, Omar Shahid, Md. Tahmeed Abdullah

    A plethora of research in the literature shows how human eye fixation pattern varies depending on different factors, including genetics, age, social functioning, cognitive functioning, and so on. Analysis of these variations in visual attention has already elicited two potential research avenues: 1) determining the physiological or psychological state of the

  2. Gourab K Patro, Abhijnan Chakraborty, Niloy Ganguly, Krishna P. Gummadi

    The (COVID-19) pandemic-induced restrictions on travel and social gatherings have prompted most conference organizers to move their events online. However, in contrast to physical conferences, virtual conferences face a challenge in efficiently scheduling talks, accounting for the availability of participants from different time-zones as well as their intere

  3. Camilo Thorne, Saber Akhondi

    We evaluate chemical patent word embeddings against known biomedical embeddings and show that they outperform the latter extrinsically and intrinsically. We also show that using contextualized embeddings can induce predictive models of reasonable performance for this domain over a relatively small gold standard.

  4. John C. Raymond, Jonathan D. Slavin, William P. Blair, Igor V. Chilingarian

    Radiative shock waves in the Cygnus Loop and other supernova remnants show different morphologies in [O III] and H{\alpha} emission. We use HST spectra and narrowband images to study the development of turbulence in the cooling region behind a shock on the west limb of the Cygnus Loop. We refine our earlier estimates of shock parameters that were based upon

  5. M. Wiedeking, M. Guttormsen, A. C. Larsen, F. Zeiser

    The Shape method, a novel approach to obtain the functional form of the $\gamma$-ray strength function ($\gamma$SF) in the absence of neutron resonance spacing data, is introduced. When used in connection with the Oslo method the slope of the Nuclear Level Density (NLD) is obtained simultaneously. The foundation of the Shape method lies in the primary $\gamm

  6. Alexander B. Balakin, Amir F. Shakirzyanov

    We consider an axionic dark matter model with a modified periodic potential for the pseudoscalar field in the framework of the axionic extension of the Einstein-aether theory. The modified potential is assumed to be equipped by the guiding function, which depends on the expansion scalar constructed as the trace of the covariant derivative of the aether veloc

  7. Depen Morwani, Harish G. Ramaswamy

    We analyze the inductive bias of gradient descent for weight normalized smooth homogeneous neural nets, when trained on exponential or cross-entropy loss. We analyse both standard weight normalization (SWN) and exponential weight normalization (EWN), and show that the gradient flow path with EWN is equivalent to gradient flow on standard networks with an ada

  8. Xiang Ling, Lingfei Wu, Saizhuo Wang, Gaoning Pan

    Code retrieval is to find the code snippet from a large corpus of source code repositories that highly matches the query of natural language description. Recent work mainly uses natural language processing techniques to process both query texts (i.e., human natural language) and code snippets (i.e., machine programming language), however neglecting the deep

  9. Yongqi Zhang, Quanming Yao, Lei Chen

    Negative sampling, which samples negative triplets from non-observed ones in knowledge graph (KG), is an essential step in KG embedding. Recently, generative adversarial network (GAN), has been introduced in negative sampling. By sampling negative triplets with large gradients, these methods avoid the problem of vanishing gradient and thus obtain better perf

  10. Vaidehi S. Paliya, A. Domínguez, C. Cabello, N. Cardiel

    One of the major challenges in studying the cosmic evolution of relativistic jets is the identification of the high-redshift ($z>3$) BL Lacertae objects, a class of jetted active galactic nuclei characterized by their quasi-featureless optical spectra. Here we report the identification of the first $\gamma$-ray emitting BL Lac object, 4FGL~J1219.0+3653 (J121

  11. Dimiter Dobrev

    We will reduce the task of creating AI to the task of finding an appropriate language for description of the world. This will not be a programing language because programing languages describe only computable functions, while our language will describe a somewhat broader class of functions. Another specificity of this language will be that the description wi

  12. Ai-Jun Ma, Wen-Fei Wang

    We study the contributions of the kaon pair originating from the resonance $\rho(770)$ for the three-body decays $B \to D K\bar{K}$ by employing the perturbative QCD approach. According to the predictions in this work, the contributions from the intermediate state $\rho(770)^0 $ are relatively small for the three-body decays such as $B^0 \to \bar{D}^0 K^+ K^

  13. Masahiro Kato, Zhenghang Cui, Yoshihiro Fukuhara

    This paper proposes a classification framework with a rejection option to mitigate the performance deterioration caused by adversarial examples. While recent machine learning algorithms achieve high prediction performance, they are empirically vulnerable to adversarial examples, which are slightly perturbed data samples that are wrongly classified. In real-w

  14. Raul Ciancarella, Francesco Pannarale, Andrea Addazi, Antonino Marciano

    We inspect the possibility that neutron star interiors are a mixture of ordinary matter and mirror dark matter. This is a scenario that can be naturally envisaged according to well studied accretion mechanisms, including the Bondi-Hoyle one. We show that the inclusion of mirror dark matter in neutron star models lowers the maximum neutron star mass for a giv

  15. A. Brudnyi

    We study exponential factorization of invertible matrices over unital complex Banach algebras. In particular, we prove that every invertible matrix with entries in the algebra of holomorphic functions on a closed bordered Riemann surface can be written as a product of two exponents of matrices over this algebra. Our result extends similar results proved earl

  16. Ginger Egberts, Fred Vermolen, Paul van Zuijlen

    We consider a one-dimensional morphoelastic model describing post-burn scar contraction. This model describes the displacement of the dermal layer of the skin and the development of the effective Eulerian strain in the tissue. Besides these components, the model also contains components that play a major role in skin repair after trauma. These components are

  17. Anna Cima, Armengol Gasull, Víctor Mañosa, Francesc Mañosas

    We describe the global dynamics of some pointwise periodic piecewise linear maps in the plane that exhibit interesting dynamic features. For each of these maps we find a first integral. For these integrals the set of values are discrete, thus quantized. Furthermore, the level sets are bounded sets whose interior is formed by a finite number of open tiles of

  18. Shota Inagaki, Shiu Mochiyama, Takashi Hikihara

    In this study, electric power is processed using the logic operation method and the error correction algorithms to meet load demand. Electric power was treated as physically flow through the distribution network, which was governed by circuit configuration and efficiency. The hardware required to digitize or packetize electric power, which is called power pa

  19. Yana Qin, Danye Wu, Zhiwei Xu, Jie Tian

    To enhance the quality and speed of data processing and protect the privacy and security of the data, edge computing has been extensively applied to support data-intensive intelligent processing services at edge. Among these data-intensive services, ensemble learning-based services can in natural leverage the distributed computation and storage resources at

  20. Anindita Banerjee, Deepika Aggarwal, Ankush Sharma, Ganesh Yadav

    Quantum random number generators are becoming mandatory in a demanding technology world of high performing learning algorithms and security guidelines. Our implementation based on principles of quantum mechanics enable us to achieve the required randomness. We have generated high-quality quantum random numbers from a weak coherent source at telecommunication

  21. Ginger Egberts, Fred Vermolen, Paul van Zuijlen

    To deal with permanent deformations and residual stresses, we consider a morphoelastic model for the scar formation as the result of wound healing after a skin trauma. Next to the mechanical components such as strain and displacements, the model accounts for biological constituents such as the concentration of signaling molecules, the cellular densities of f

  22. Antonis Kakas, Loizos Michael

    This paper presents Abduction and Argumentation as two principled forms for reasoning, and fleshes out the fundamental role that they can play within Machine Learning. It reviews the state-of-the-art work over the past few decades on the link of these two reasoning forms with machine learning work, and from this it elaborates on how the explanation-generatin

  23. Jiyanglin Li, Tao Li

    With regard to a three-step estimation procedure, proposed without theoretical discussion by Li and You in Journal of Applied Statistics and Management, for a nonparametric regression model with time-varying regression function, local stationary regressors and time-varying AR(p) (tvAR(p)) error process , we established all necessary asymptotic properties for

  24. Sujunjie Sun, Guopeng Zhang, Haibo Mei, Kezhi Wang

    In Unmanned Aerial Vehicle (UAV)-enabled mobile edge computing (MEC) systems, UAVs can carry edge servers to help ground user equipment (UEs) offloading their computing tasks to the UAVs for execution. This paper aims to minimize the total time required for the UAVs to complete the offloaded tasks, while optimizing the three-dimensional (3D) deployment of UA

  25. Armengol Gasull, Luis Hernández-Corbato, Francisco R. Ruiz del Portal

    We construct two planar homeomorphisms $f$ and $g$ for which the origin is a globally asymptotically stable fixed point whereas for $f \circ g$ and $g \circ f$ the origin is a global repeller. Furthermore, the origin remains a global repeller for the iterated function system generated by $f$ and $g$ where each of the maps appears with a certain probability.

  26. Zuhair Al-Johar

    This article examines the notion of invariance under different kinds of permutations in a milieu of a theory of classes and sets, as a semantic motivation for Quine's new foundations "NF". The approach largely depends on interpreting a finite axiomatization of NF beginning from the least restrictions on permutations and then gradually upgrading those restric

  27. Christoph Haase, Jakub Różycki

    We show that the existential fragment of B\"uchi arithmetic is strictly less expressive than full B\"uchi arithmetic of any base, and moreover establish that its $\Sigma_2$-fragment is already expressively complete. Furthermore, we show that regular languages of polynomial growth are definable in the existential fragment of B\"uchi arithmetic.

  28. M. Osiekowicz, D. Staszczuk, K. Olkowska-Pucko, Ł. Kipczak

    The temperature effect on the Raman scattering efficiency is investigated in $\varepsilon$-GaSe and $\gamma$-InSe crystals. We found that varying the temperature over a broad range from 5 K to 350 K permits to achieve both the resonant conditions and the antiresonance behaviour in Raman scattering of the studied materials. The resonant conditions of Raman sc

  29. Liangyi Huang, Hui Rao

    Let E be a metric space. We introduce a notion of connectedness index of E, which is the Hausdor? dimension of the union of non-trivial connected components of E. We show that the connectedness index of a fractal cube E is strictly less than the Hausdor? dimension of E provided that E possesses a trivial connected component. Hence the connectedness index is

  30. Alejandro Donaire, Luigi Villani, Fanny Ficuciello, Juan Tomassini

    In this paper, we present an impedance control design for multi-variable linear and nonlinear robotic systems. The control design considers force and state feedback to improve the performance of the closed loop. Simultaneous feedback of forces and states allows the controller for an extra degree of freedom to approximate the desired impedance port behaviour.

  31. Sungho Suh, Paul Lukowicz, Yong Oh Lee

    The data imbalance problem is a frequent bottleneck in the classification performance of neural networks. In this paper, we propose a novel supervised discriminative feature generation (DFG) method for a minority class dataset. DFG is based on the modified structure of a generative adversarial network consisting of four independent networks: generator, discr

  32. Jincheng Bai, Qifan Song, Guang Cheng

    We propose a variational Bayesian (VB) procedure for high-dimensional linear model inferences with heavy tail shrinkage priors, such as student-t prior. Theoretically, we establish the consistency of the proposed VB method and prove that under the proper choice of prior specifications, the contraction rate of the VB posterior is nearly optimal. It justifies

  33. Zhuolin Ye, Viktor Holubec

    We consider absorption refrigerators consisting of simultaneously operating Carnot-type heat engine and refrigerator. Their maximum efficiency at given power (MEGP) is given by the product of MEGPs for the internal engine and refrigerator. The only subtlety of the derivation lies in the fact that the maximum cooling power of the absorption refrigerator is no

  34. Tong Niu, Semih Yavuz, Yingbo Zhou, Nitish Shirish Keskar

    Paraphrase generation has benefited extensively from recent progress in the designing of training objectives and model architectures. However, previous explorations have largely focused on supervised methods, which require a large amount of labeled data that is costly to collect. To address this drawback, we adopt a transfer learning approach and propose a t

  35. Ximing Lu, Peter West, Rowan Zellers, Ronan Le Bras

    Conditional text generation often requires lexical constraints, i.e., which words should or shouldn't be included in the output text. While the dominant recipe for conditional text generation has been large-scale pretrained language models that are finetuned on the task-specific training data, such models do not learn to follow the underlying constraints rel

  36. Quoc Phong Nguyen, Bryan Kian Hsiang Low, Patrick Jaillet

    This paper studies the problem of approximately unlearning a Bayesian model from a small subset of the training data to be erased. We frame this problem as one of minimizing the Kullback-Leibler divergence between the approximate posterior belief of model parameters after directly unlearning from erased data vs. the exact posterior belief from retraining wit

  37. Mingyang Chen, Wen Zhang, Zonggang Yuan, Yantao Jia

    Knowledge graphs (KGs) consisting of triples are always incomplete, so it's important to do Knowledge Graph Completion (KGC) by predicting missing triples. Multi-Source KG is a common situation in real KG applications which can be viewed as a set of related individual KGs where different KGs contains relations of different aspects of entities. It's intuitive

  38. Arturo Oncevay, Kervy Rivas Rojas

    Language modelling is regularly analysed at word, subword or character units, but syllables are seldom used. Syllables provide shorter sequences than characters, they can be extracted with rules, and their segmentation typically requires less specialised effort than identifying morphemes. We reconsider syllables for an open-vocabulary generation task in 20 l

  39. Mehdi Bonyani, Simindokht Jahangard, Morteza Daneshmand

    Digit, letter and word recognition for a particular script has various applications in todays commercial contexts. Nevertheless, only a limited number of relevant studies have dealt with Persian scripts. In this paper, deep neural networks are utilized through various DensNet architectures, as well as the Xception, are adopted, modified and further boosted t

  40. Zbigniew Palmowski, Daria Puchalska

    The contagion dynamics can emerge in social networks when repeated activation is allowed. An interesting example of this phenomenon is retweet cascades where users allow to re-share content posted by other people with public accounts. To model this type of behaviour we use a Hawkes self-exciting process. To do it properly though one needs to calibrate model

  41. Norman Haussmann, Martin Zang, Robin Mease, Markus Clemens

    The exposure of a human by magneto-quasistatic fields from wireless charging systems is to be determined from magnetic field measurements in near real-time. This requires a fast linear equations solver for the discrete Poisson system of the Co-Simulation Scalar Potential Finite Difference (Co-Sim. SPFD) scheme. Here, the use of the AmgX library on NVIDIA GPU

  42. Benedek Rozemberczki, Peter Englert, Amol Kapoor, Martin Blais

    In this work we propose Pathfinder Discovery Networks (PDNs), a method for jointly learning a message passing graph over a multiplex network with a downstream semi-supervised model. PDNs inductively learn an aggregated weight for each edge, optimized to produce the best outcome for the downstream learning task. PDNs are a generalization of attention mechanis

  43. Fardin Ghorbani, Javad Shabanpour, Sepideh Monjezi, Hossein Soleimani

    In the quest to realize a comprehensive EEG signal processing framework, in this paper, we demonstrate a toolbox and graphic user interface, EEGsig, for the full process of EEG signals. Our goal is to provide a comprehensive suite, free and open-source framework for EEG signal processing where the users especially physicians who do not have programming exper

  44. Gexin Huang, Jiawen Liang, Ke Liu, Chang Cai

    Electromagnetic source imaging (ESI) requires solving a highly ill-posed inverse problem. To seek a unique solution, traditional ESI methods impose various forms of priors that may not accurately reflect the actual source properties, which may hinder their broad applications. To overcome this limitation, in this paper a novel data-synthesized spatio-temporal

  45. Yu Tian, Gaofeng Pan, Mohamed-Slim Alouini

    In this paper, a dual-hop cooperative satellite-unmanned aerial vehicle (UAV) communication system including a satellite (S), a group of cluster headers (CHs), which are respectively with a group of uniformly distributed UAVs, is considered. Specifically, these CHs serve as aerial decode-and-forward relays to forward the information transmitted by S to UAVs.

  46. Ganesh C. Paul, SK Firoz Islam, Paramita Dutta, Arijit Saha

    We theoretically investigate the features of Ruderman-Kittel-Kasuya-Yosida (RKKY) exchange interaction between two magnetic impurities, mediated by the interfacial bound states inside a domain wall (DW). The latter separates the two regions with oppositely signed inversion symmetry broken terms in graphene and Weyl semimetal. The DW is modelled by a smooth q

  47. Jun Yan, Mrigank Raman, Aaron Chan, Tianyu Zhang

    Recently, knowledge graph (KG) augmented models have achieved noteworthy success on various commonsense reasoning tasks. However, KG edge (fact) sparsity and noisy edge extraction/generation often hinder models from obtaining useful knowledge to reason over. To address these issues, we propose a new KG-augmented model: Hybrid Graph Network (HGN). Unlike prio

  48. Mrigank Raman, Aaron Chan, Siddhant Agarwal, Peifeng Wang

    Knowledge graphs (KGs) have helped neural models improve performance on various knowledge-intensive tasks, like question answering and item recommendation. By using attention over the KG, such KG-augmented models can also "explain" which KG information was most relevant for making a given prediction. In this paper, we question whether these models are really

  49. Zein Shaheen, Gerhard Wohlgenannt, Erwin Filtz

    Large multi-label text classification is a challenging Natural Language Processing (NLP) problem that is concerned with text classification for datasets with thousands of labels. We tackle this problem in the legal domain, where datasets, such as JRC-Acquis and EURLEX57K labeled with the EuroVoc vocabulary were created within the legal information systems of

  50. Ahmed Touati, Pascal Vincent

    We study episodic reinforcement learning in non-stationary linear (a.k.a. low-rank) Markov Decision Processes (MDPs), i.e, both the reward and transition kernel are linear with respect to a given feature map and are allowed to evolve either slowly or abruptly over time. For this problem setting, we propose OPT-WLSVI an optimistic model-free algorithm based o

  51. Suresh Nambi, Salim Ullah, Aditya Lohana, Siva Satyendra Sahoo

    The recent advances in machine learning, in general, and Artificial Neural Networks (ANN), in particular, has made smart embedded systems an attractive option for a larger number of application areas. However, the high computational complexity, memory footprints, and energy requirements of machine learning models hinder their deployment on resource-constrain

  52. Yongchang Hao, Shilin He, Wenxiang Jiao, Zhaopeng Tu

    Non-Autoregressive machine Translation (NAT) models have demonstrated significant inference speedup but suffer from inferior translation accuracy. The common practice to tackle the problem is transferring the Autoregressive machine Translation (AT) knowledge to NAT models, e.g., with knowledge distillation. In this work, we hypothesize and empirically verify

  53. Syed Muhammad Kazim, Ahmad Farooq, Junaid ur Rehman, Hyundong Shin

    Several Bayesian estimation based heuristics have been developed to perform quantum state tomography (QST). Their ability to quantify uncertainties using region estimators and include a priori knowledge of the experimentalists makes this family of methods an attractive choice for QST. However, specialized techniques for pure states do not work well for mixed

  54. Runbing Zheng, Vince Lyzinski, Carey E. Priebe, Minh Tang

    Given a network and a subset of interesting vertices whose identities are only partially known, the vertex nomination problem seeks to rank the remaining vertices in such a way that the interesting vertices are ranked at the top of the list. An important variant of this problem is vertex nomination in the multi-graphs setting. Given two graphs $G_1, G_2$ wit

  55. Mou-Cheng Xu

    Brain shift makes the pre-operative MRI navigation highly inaccurate hence the intraoperative modalities are adopted in surgical theatre. Due to the excellent economic and portability merits, the Ultrasound imaging is used at our collaborating hospital, Charing Cross Hospital, Imperial College London, UK. However, it is found that intraoperative diagnosis on

  56. Kyungjae Lee, Hongjun Yang, Sungbin Lim, Songhwai Oh

    In this paper, we consider stochastic multi-armed bandits (MABs) with heavy-tailed rewards, whose $p$-th moment is bounded by a constant $\nu_{p}$ for $1<p\leq2$. First, we propose a novel robust estimator which does not require $\nu_{p}$ as prior information, while other existing robust estimators demand prior knowledge about $\nu_{p}$. We show that an erro

  57. Jiajin Li, Caihua Chen, Anthony Man-Cho So

    Wasserstein \textbf{D}istributionally \textbf{R}obust \textbf{O}ptimization (DRO) is concerned with finding decisions that perform well on data that are drawn from the worst-case probability distribution within a Wasserstein ball centered at a certain nominal distribution. In recent years, it has been shown that various DRO formulations of learning models ad

  58. Xisen Jin, Francesco Barbieri, Brendan Kennedy, Aida Mostafazadeh Davani

    Fine-tuned language models have been shown to exhibit biases against protected groups in a host of modeling tasks such as text classification and coreference resolution. Previous works focus on detecting these biases, reducing bias in data representations, and using auxiliary training objectives to mitigate bias during fine-tuning. Although these techniques

  59. Javlon Rayimbaev, Pulat Tadjimuratov, Ahmadjon Abdujabbarov, Bobomurat Ahmedov

    In the work, we have presented detailed analyses of the event horizon and curvature properties of spacetime around a regular black hole in modified gravity so-called regular MOG black hole. The motion of neutral, electrically charged and magnetized particles with magnetic dipole moment, in the close environment of the regular MOG black hole immersed in an ex

  60. Ainur Zhaikhan, Mustafa A. Kishk, Hesham ElSawy, Mohamed-Slim Alouini

    The upcoming Internet of things (IoT) is foreseen to encompass massive numbers of connected devices, smart objects, and cyber-physical systems. Due to the large-scale and massive deployment of devices, it is deemed infeasible to safeguard 100% of the devices with state-of-the-art security countermeasures. Hence, large-scale IoT has inevitable loopholes for n

  61. Syuan-Hao Sie, Jye-Luen Lee, Yi-Ren Chen, Chih-Cheng Lu

    Convolutional neural networks (CNNs) play a key role in deep learning applications. However, the large storage overheads and the substantial computation cost of CNNs are problematic in hardware accelerators. Computing-in-memory (CIM) architecture has demonstrated great potential to effectively compute large-scale matrix-vector multiplication. However, the in

  62. Giovanni Lapenta, Jean Berchem, Mostafa El Alaoui, Raymond Walker

    Earth's magnetotail is an excellent laboratory to study the interplay of reconnection and turbulence in determining electron energization. The process of formation of a power law tail during turbulent reconnection is a documented fact still in need of a comprehensive explanation. We conduct a massively parallel particle in cell 3D simulation and use enhanced

  63. Soufiane Hayou, Eugenio Clerico, Bobby He, George Deligiannidis

    Deep ResNet architectures have achieved state of the art performance on many tasks. While they solve the problem of gradient vanishing, they might suffer from gradient exploding as the depth becomes large (Yang et al. 2017). Moreover, recent results have shown that ResNet might lose expressivity as the depth goes to infinity (Yang et al. 2017, Hayou et al. 2

  64. Benjamin Muller, Antonis Anastasopoulos, Benoît Sagot, Djamé Seddah

    Transfer learning based on pretraining language models on a large amount of raw data has become a new norm to reach state-of-the-art performance in NLP. Still, it remains unclear how this approach should be applied for unseen languages that are not covered by any available large-scale multilingual language model and for which only a small amount of raw data

  65. Bishal Santra, Potnuru Anusha, Pawan Goyal

    Generative models for dialog systems have gained much interest because of the recent success of RNN and Transformer based models in tasks like question answering and summarization. Although the task of dialog response generation is generally seen as a sequence-to-sequence (Seq2Seq) problem, researchers in the past have found it challenging to train dialog sy

  66. Bill Yuchen Lin, Haitian Sun, Bhuwan Dhingra, Manzil Zaheer

    Current commonsense reasoning research focuses on developing models that use commonsense knowledge to answer multiple-choice questions. However, systems designed to answer multiple-choice questions may not be useful in applications that do not provide a small list of candidate answers to choose from. As a step towards making commonsense reasoning research mo

  67. William McCorkindale, Carl Poelking, Alpha A. Lee

    Predicting bioactivity and physical properties of molecules is a longstanding challenge in drug design. Most approaches use molecular descriptors based on a 2D representation of molecules as a graph of atoms and bonds, abstracting away the molecular shape. A difficulty in accounting for 3D shape is in designing molecular descriptors can precisely capture mol

  68. A. El Ichi, K. Jbilou, R. Sadaka

    In the present paper, we introduce new tensor Krylov subspace methods for solving linear tensor equations. The proposed methods use the well known T-product for tensors and tensor subspaces related to tube fibers. We introduce some new tensor products and the related algebraic properties. These new products will enable us to develop third-order the tensor tu

  69. Mohsen Kian, Yuki Seo

    Employing the notion of operator log-convexity, we study joint concavity$/$ convexity of multivariable operator functions: $(A,B)\mapsto F(A,B)=h\left[ \Phi(f(A))\ \sigma\ \Psi(g(B))\right]$, where $\Phi$ and $\Psi$ are positive linear maps and $\sigma$ is an operator mean. As applications, we prove jointly concavity$/$convexity of matrix trace functions $\T

  70. Pierre Cardaliaguet, Benjamin Seeger

    We obtain space-time H\"older regularity estimates for solutions of first- and second-order Hamilton-Jacobi equations perturbed with an additive stochastic forcing term. The bounds depend only on the growth of the Hamiltonian in the gradient and on the regularity of the stochastic coefficients, in a way that is invariant with respect to a hyperbolic scaling.

  71. Shih-Ting Lin, Ashish Sabharwal, Tushar Khot

    We present ReadOnce Transformers, an approach to convert a transformer-based model into one that can build an information-capturing, task-independent, and compressed representation of text. The resulting representation is reusable across different examples and tasks, thereby requiring a document shared across many examples or tasks to only be \emph{read once

  72. V. Jeganathan, K. Alba, R. Ostilla-Mónico

    Taylor-Couette (TC) flow, the flow between two independently rotating and co-axial cylinders is commonly used as a canonical model for shear flows. Unlike plane Couette, pinned secondary flows can be found in TC flow. These are known as Taylor rolls and drastically affect the flow behaviour. We study the possibility of modifying these secondary structures us

  73. Radhika Dua, Sai Srinivas Kancheti, Vineeth N Balasubramanian

    Visual Question Answering is a multi-modal task that aims to measure high-level visual understanding. Contemporary VQA models are restrictive in the sense that answers are obtained via classification over a limited vocabulary (in the case of open-ended VQA), or via classification over a set of multiple-choice-type answers. In this work, we present a complete

  74. Daniel Flamm, Daniel Günther Grossmann, Michael Jenne, Felix Zimmermann

    The remarkable temporal properties of ultra-short pulsed lasers in combination with novel beam shaping concepts enable the development of completely new material processing strategies. We demonstrate the benefit of employing focus distributions being tailored in all three spatial dimensions. As example advanced Bessel-like beam profiles, 3D-beam splitting co

  75. Shiyang Li, Semih Yavuz, Kazuma Hashimoto, Jia Li

    Dialogue state trackers have made significant progress on benchmark datasets, but their generalization capability to novel and realistic scenarios beyond the held-out conversations is less understood. We propose controllable counterfactuals (CoCo) to bridge this gap and evaluate dialogue state tracking (DST) models on novel scenarios, i.e., would the system

  76. Pouya Bakhti, Meshkat Rajaee

    We investigate the potential of the next generation long-baseline neutrino experiments DUNE and T2HK as well as the upcoming reactor experiment JUNO to constrain Non-Standard Interaction (NSI) parameters. JUNO is going to provide the most precise measurements of solar neutrino oscillation parameters as well as determining the neutrino mass ordering. We study

  77. Ethan A. Chi, Julian Salazar, Katrin Kirchhoff

    Non-autoregressive models greatly improve decoding speed over typical sequence-to-sequence models, but suffer from degraded performance. Infilling and iterative refinement models make up some of this gap by editing the outputs of a non-autoregressive model, but are constrained in the edits that they can make. We propose iterative realignment, where refinemen

  78. Neil Walton

    We evaluate the ability of temporal difference learning to track the reward function of a policy as it changes over time. Our results apply a new adiabatic theorem that bounds the mixing time of time-inhomogeneous Markov chains. We derive finite-time bounds for tabular temporal difference learning and $Q$-learning when the policy used for training changes in

  79. Abdelhakim Hannousse, Salima Yahiouche

    In this paper, we present a general scheme for building reproducible and extensible datasets for website phishing detection. The aim is to (1) enable comparison of systems using different features, (2) overtake the short-lived nature of phishing websites, and (3) keep track of the evolution of phishing tactics. For experimenting the proposed scheme, we start

  80. Ben Li, Fabian Mussnig

    We introduce a class of functional analogs of the symmetric difference metric on the space of coercive convex functions on $\mathbb{R}^n$ with full-dimensional domain. We show that convergence with respect to these metrics is equivalent to epi-convergence. Furthermore, we give a full classification of all isometries with respect to some of the new metrics. M

  81. Séverin Philip

    In this paper we study the field of definition of abelian subvarieties $B\subset A_{\overline{K}}$ for an abelian variety $A$ over a field $K$ of characteristic $0$. We show that, provided that no isotypic component of $A_{\overline{K}}$ is simple, there are infinitely many abelian subvarieties of $A_{\overline{K}}$ with field of definition $K_A$, the field

  82. Sahisnu Mazumder, Oriana Riva

    AI assistants can now carry out tasks for users by directly interacting with website UIs. Current semantic parsing and slot-filling techniques cannot flexibly adapt to many different websites without being constantly re-trained. We propose FLIN, a natural language interface for web navigation that maps user commands to concept-level actions (rather than low-

  83. Jakub Slavík

    We establish the large deviations principle (LDP) and the moderate deviations principle (MDP) and an almost sure version of the central limit theorem (CLT) for the stochastic 3D viscous primitive equations driven by a multiplicative white noise allowing dependence on spatial gradient of solutions with initial data in $H^2$. The LDP is established using the w

  84. Nicole Mücke

    Stochastic gradient descent (SGD) provides a simple and efficient way to solve a broad range of machine learning problems. Here, we focus on distribution regression (DR), involving two stages of sampling: Firstly, we regress from probability measures to real-valued responses. Secondly, we sample bags from these distributions for utilizing them to solve the o

  85. Muhammad Yousefnezhad, Alessandro Selvitella, Daoqiang Zhang, Andrew J. Greenshaw

    Multi-voxel pattern analysis (MVPA) learns predictive models from task-based functional magnetic resonance imaging (fMRI) data, for distinguishing when subjects are performing different cognitive tasks -- e.g., watching movies or making decisions. MVPA works best with a well-designed feature set and an adequate sample size. However, most fMRI datasets are no

  86. Amit Anand, Bikash K. Behera, Prasanta K. Panigrahi

    Diners dilemma is one of the most interesting problems in both economic and game theories. Here, we solve this problem for n (number of players) =4 with quantum rules and we are able to remove the dilemma of diners between the Pareto optimal and Nash equilibrium points of the game. We find the quantum strategy that gives maximum payoff for each diner without

  87. Amirreza Silani, Michele Cucuzzella, Jacquelien M. A. Scherpen, Mohammad Javad Yazdanpanah

    Motivated by the inadequacy of the existing control strategies for power systems affected by time-varying uncontrolled power injections such as loads and the increasingly widespread renewable energy sources, this paper proposes two control schemes based on the well-known output regulation control methodology. The first one is designed based on the classical

  88. Sina Sajjadi, Alireza Hashemi, Fakhteh Ghanbarnejad

    Non-pharmaceutical measures such as social distancing, can play an important role to control an epidemic in the absence of vaccinations. In this paper, we study the impact of social distancing on epidemics for which it is executable. We use a mathematical model combining human mobility and disease spreading. For the mobility dynamics, we design an agent base

  89. Junyong Zhang

    We construct the Schwartz kernel of resolvent and spectral measure for Schr\"odinger operators on the flat Euclidean cone $(X,g)$, where $X=C(\mathbb{S}_\sigma^1)=(0,\infty)\times \mathbb{S}_\sigma^1$ is a product cone over the circle, $\mathbb{S}_\sigma^1=\R/2\pi \sigma\Z$, with radius $\sigma>0$ and the metric $g=dr^2+r^2 d\theta^2$. As products, we prove

  90. Fuyu Lv, Mengxue Li, Tonglei Guo, Changlong Yu

    Deep learning-based sequential recommender systems have recently attracted increasing attention from both academia and industry. Most of industrial Embedding-Based Retrieval (EBR) system for recommendation share the similar ideas with sequential recommenders. Among them, how to comprehensively capture sequential user interest is a fundamental problem. Howeve

  91. Alexander R. Fabbri, Simeng Han, Haoyuan Li, Haoran Li

    Models pretrained with self-supervised objectives on large text corpora achieve state-of-the-art performance on English text summarization tasks. However, these models are typically fine-tuned on hundreds of thousands of data points, an infeasible requirement when applying summarization to new, niche domains. In this work, we introduce a novel and generaliza

  92. Muhammad Sufyan, Hamayun Farooq, Imran Akhtar, Zafar Bangash

    Proper orthogonal decomposition (POD) is often employed in developing reduced-order models (ROM) in fluid flows for design, control, and optimization. Contrary to the usual practice where velocity field is the focus, we apply the POD analysis on the pressure field data obtained from numerical simulations of the flow past stationary and oscillating cylinders.

  93. Saadia Gabriel, Asli Celikyilmaz, Rahul Jha, Yejin Choi

    While neural language models can generate text with remarkable fluency and coherence, controlling for factual correctness in generation remains an open research question. This major discrepancy between the surface-level fluency and the content-level correctness of neural generation has motivated a new line of research that seeks automatic metrics for evaluat

  94. Georgia Papacharalampous, Hristos Tyralis, Simon Michael Papalexiou, Andreas Langousis

    Hydroclimatic time series analysis focuses on a few feature types (e.g., autocorrelations, trends, extremes), which describe a small portion of the entire information content of the observations. Aiming to exploit a larger part of the available information and, thus, to deliver more reliable results (e.g., in hydroclimatic time series clustering contexts), h

  95. Yi Xing, Jie Zheng, Jindong Li, Yixin Cao

    We perform combined X-ray tomography and shear force measurements on a cyclically sheared granular system with highly transient behaviors, and obtain the evolution of microscopic structures and the macroscopic shear force during the shear cycle. We explain the macroscopic behaviors of the system based on microscopic processes, including the particle level st

  96. Liunian Harold Li, Haoxuan You, Zhecan Wang, Alireza Zareian

    Pre-trained contextual vision-and-language (V&L) models have achieved impressive performance on various benchmarks. However, existing models require a large amount of parallel image-caption data for pre-training. Such data are costly to collect and require cumbersome curation. Inspired by unsupervised machine translation, we investigate if a strong V&L repre

  97. Timothée Bénard

    Let $G$ be a connected simple real Lie group, $\Lambda_{0}\subseteq G$ a lattice and $\Lambda \unlhd \Lambda_{0}$ a normal subgroup such that $\Lambda_{0}/\Lambda\simeq \mathbb{Z}^d$. We study the drift of a random walk on the $\mathbb{Z}^d$-cover $\Lambda\backslash G$ of the finite volume homogeneous space $\Lambda_{0}\backslash G$. This walk is defined by

  98. Xian Li, Changhan Wang, Yun Tang, Chau Tran

    We present a simple yet effective approach to build multilingual speech-to-text (ST) translation by efficient transfer learning from pretrained speech encoder and text decoder. Our key finding is that a minimalistic LNA (LayerNorm and Attention) finetuning can achieve zero-shot crosslingual and cross-modality transfer ability by only finetuning less than 10%

  99. Haoyu Zhang, Dingkun Long, Guangwei Xu, Pengjun Xie

    Keyphrase extraction (KE) aims to summarize a set of phrases that accurately express a concept or a topic covered in a given document. Recently, Sequence-to-Sequence (Seq2Seq) based generative framework is widely used in KE task, and it has obtained competitive performance on various benchmarks. The main challenges of Seq2Seq methods lie in acquiring informa

  100. Amane Sugiyama, Naoki Yoshinaga

    Although many context-aware neural machine translation models have been proposed to incorporate contexts in translation, most of those models are trained end-to-end on parallel documents aligned in sentence-level. Because only a few domains (and language pairs) have such document-level parallel data, we cannot perform accurate context-aware translation in mo