https://arxiv.org/api/z+GeFugfesBBTOQe1SL5h3z+aK4 arXiv Query: search_query=&id_list=2401.00401,2401.00402,2401.00403,2401.00404,2401.00405,2401.00406,2401.00407,2401.00408,2401.00409,2401.00410,2401.00411,2401.00412,2401.00413,2401.00414,2401.00415,2401.00416,2401.00417,2401.00418,2401.00419,2401.00420,2401.00421,2401.00422,2401.00423,2401.00424,2401.00425,2401.00426,2401.00427,2401.00428,2401.00429,2401.00430,2401.00431,2401.00432,2401.00433,2401.00434,2401.00435,2401.00436,2401.00437,2401.00438,2401.00439,2401.00440,2401.00441,2401.00442,2401.00443,2401.00444,2401.00445,2401.00446,2401.00447,2401.00448,2401.00449,2401.00450,2401.00451,2401.00452,2401.00453,2401.00454,2401.00455,2401.00456,2401.00457,2401.00458,2401.00459,2401.00460,2401.00461,2401.00462,2401.00463,2401.00464,2401.00465,2401.00466,2401.00467,2401.00468,2401.00469,2401.00470,2401.00471,2401.00472,2401.00473,2401.00474,2401.00475,2401.00476,2401.00477,2401.00478,2401.00479,2401.00480,2401.00481,2401.00482,2401.00483,2401.00484,2401.00485,2401.00486,2401.00487,2401.00488,2401.00489,2401.00490,2401.00491,2401.00492,2401.00493,2401.00494,2401.00495,2401.00496,2401.00497,2401.00498,2401.00499,2401.00500&start=0&max_results=100 2026-07-02T22:41:38Z 100 100 0 http://arxiv.org/abs/2401.00402v1 3D Multi-system Bayesian Calibration with Energy Conservation to Study Rapidity-dependent Dynamics of Nuclear Collisions 2023-12-31T05:33:57Z Considerable information about the early-stage dynamics of heavy-ion collisions is encoded in the rapidity dependence of measurements. To leverage the large amount of experimental data, we perform a systematic analysis using three-dimensional hydrodynamic simulations of multiple collision systems -- large and small, symmetric and asymmetric. Specifically, we perform fully 3D multi-stage hydrodynamic simulations initialized by a parameterized model for rapidity-dependent energy deposition, which we calibrate on the hadron multiplicity and anisotropic flow coefficients. We utilize Bayesian inference to constrain properties of the early- and late- time dynamics of the system, and highlight the impact of enforcing global energy conservation in our 3D model. 2023-12-31T05:33:57Z Andi Mankolli The JETSCAPE Collaboration Aaron Angerami The JETSCAPE Collaboration Ritu Arora The JETSCAPE Collaboration Steffen Bass The JETSCAPE Collaboration Shanshan Cao The JETSCAPE Collaboration Yi Chen The JETSCAPE Collaboration Lipei Du The JETSCAPE Collaboration Raymond Ehlers The JETSCAPE Collaboration Hannah Elfner The JETSCAPE Collaboration Wenkai Fan The JETSCAPE Collaboration Rainer J. Fries The JETSCAPE Collaboration Charles Gale The JETSCAPE Collaboration Yayun He The JETSCAPE Collaboration Ulrich Heinz The JETSCAPE Collaboration Barbara Jacak The JETSCAPE Collaboration Peter Jacobs The JETSCAPE Collaboration Sangyong Jeon The JETSCAPE Collaboration Yi Ji The JETSCAPE Collaboration Lauren Kasper The JETSCAPE Collaboration Michael Kordell The JETSCAPE Collaboration Amit Kumar The JETSCAPE Collaboration R. Kunnawalkam-Elayavalli The JETSCAPE Collaboration Joseph Latessa The JETSCAPE Collaboration Sook H. Lee The JETSCAPE Collaboration Yen-Jie Lee The JETSCAPE Collaboration Dananjaya Liyanage The JETSCAPE Collaboration Matt Luzum The JETSCAPE Collaboration Abhijit Majumder The JETSCAPE Collaboration Simon Mak The JETSCAPE Collaboration Christal Martin The JETSCAPE Collaboration Haydar Mehryar The JETSCAPE Collaboration Tanner Mengel The JETSCAPE Collaboration James Mulligan The JETSCAPE Collaboration Christine Nattrass The JETSCAPE Collaboration Jean-Francois Paquet The JETSCAPE Collaboration Cameron Parker The JETSCAPE Collaboration Joern H. Putschke The JETSCAPE Collaboration Gunther Roland The JETSCAPE Collaboration Bjoern Schenke The JETSCAPE Collaboration Loren Schwiebert The JETSCAPE Collaboration Arjun Sengupta The JETSCAPE Collaboration Chun Shen The JETSCAPE Collaboration Chathuranga Sirimanna The JETSCAPE Collaboration Ron A. Soltz The JETSCAPE Collaboration Ismail Soudi The JETSCAPE Collaboration Michael Strickland The JETSCAPE Collaboration Yasuki Tachibana The JETSCAPE Collaboration Julia Velkovska The JETSCAPE Collaboration Gojko Vujanovic The JETSCAPE Collaboration Xin-Nian Wang The JETSCAPE Collaboration Wenbin Zhao The JETSCAPE Collaboration http://arxiv.org/abs/2401.00418v1 Bounds on the minimum distance of locally recoverable codes 2023-12-31T07:47:16Z We consider locally recoverable codes (LRCs) and aim to determine the smallest possible length $n=n_q(k,d,r)$ of a linear $[n,k,d]_q$-code with locality $r$. For $k\le 7$ we exactly determine all values of $n_2(k,d,2)$ and for $k\le 6$ we exactly determine all values of $n_2(k,d,1)$. For the ternary field we also state a few numerical results. As a general result we prove that $n_q(k,d,r)$ equals the Griesmer bound if the minimum Hamming distance $d$ is sufficiently large and all other parameters are fixed. 2023-12-31T07:47:16Z 23 pages, 3 tables Sascha Kurz http://arxiv.org/abs/2401.00420v2 SynCDR : Training Cross Domain Retrieval Models with Synthetic Data 2024-03-19T16:56:53Z In cross-domain retrieval, a model is required to identify images from the same semantic category across two visual domains. For instance, given a sketch of an object, a model needs to retrieve a real image of it from an online store's catalog. A standard approach for such a problem is learning a feature space of images where Euclidean distances reflect similarity. Even without human annotations, which may be expensive to acquire, prior methods function reasonably well using unlabeled images for training. Our problem constraint takes this further to scenarios where the two domains do not necessarily share any common categories in training data. This can occur when the two domains in question come from different versions of some biometric sensor recording identities of different people. We posit a simple solution, which is to generate synthetic data to fill in these missing category examples across domains. This, we do via category preserving translation of images from one visual domain to another. We compare approaches specifically trained for this translation for a pair of domains, as well as those that can use large-scale pre-trained text-to-image diffusion models via prompts, and find that the latter can generate better replacement synthetic data, leading to more accurate cross-domain retrieval models. Our best SynCDR model can outperform prior art by up to 15\%. Code for our work is available at https://github.com/samarth4149/SynCDR . 2023-12-31T08:06:53Z Pre-print Samarth Mishra Carlos D. Castillo Hongcheng Wang Kate Saenko Venkatesh Saligrama http://arxiv.org/abs/2401.00421v1 From Text to Pixels: A Context-Aware Semantic Synergy Solution for Infrared and Visible Image Fusion 2023-12-31T08:13:47Z With the rapid progression of deep learning technologies, multi-modality image fusion has become increasingly prevalent in object detection tasks. Despite its popularity, the inherent disparities in how different sources depict scene content make fusion a challenging problem. Current fusion methodologies identify shared characteristics between the two modalities and integrate them within this shared domain using either iterative optimization or deep learning architectures, which often neglect the intricate semantic relationships between modalities, resulting in a superficial understanding of inter-modal connections and, consequently, suboptimal fusion outcomes. To address this, we introduce a text-guided multi-modality image fusion method that leverages the high-level semantics from textual descriptions to integrate semantics from infrared and visible images. This method capitalizes on the complementary characteristics of diverse modalities, bolstering both the accuracy and robustness of object detection. The codebook is utilized to enhance a streamlined and concise depiction of the fused intra- and inter-domain dynamics, fine-tuned for optimal performance in detection tasks. We present a bilevel optimization strategy that establishes a nexus between the joint problem of fusion and detection, optimizing both processes concurrently. Furthermore, we introduce the first dataset of paired infrared and visible images accompanied by text prompts, paving the way for future research. Extensive experiments on several datasets demonstrate that our method not only produces visually superior fusion results but also achieves a higher detection mAP over existing methods, achieving state-of-the-art results. 2023-12-31T08:13:47Z 10 pages, 12 figures, 3 tables, conference Xingyuan Li Yang Zou Jinyuan Liu Zhiying Jiang Long Ma Xin Fan Risheng Liu http://arxiv.org/abs/2401.00424v1 SDIF-DA: A Shallow-to-Deep Interaction Framework with Data Augmentation for Multi-modal Intent Detection 2023-12-31T08:33:37Z Multi-modal intent detection aims to utilize various modalities to understand the user's intentions, which is essential for the deployment of dialogue systems in real-world scenarios. The two core challenges for multi-modal intent detection are (1) how to effectively align and fuse different features of modalities and (2) the limited labeled multi-modal intent training data. In this work, we introduce a shallow-to-deep interaction framework with data augmentation (SDIF-DA) to address the above challenges. Firstly, SDIF-DA leverages a shallow-to-deep interaction module to progressively and effectively align and fuse features across text, video, and audio modalities. Secondly, we propose a ChatGPT-based data augmentation approach to automatically augment sufficient training data. Experimental results demonstrate that SDIF-DA can effectively align and fuse multi-modal features by achieving state-of-the-art performance. In addition, extensive analyses show that the introduced data augmentation approach can successfully distill knowledge from the large language model. 2023-12-31T08:33:37Z Accepted by ICASSP 2024 Shijue Huang Libo Qin Bingbing Wang Geng Tu Ruifeng Xu http://arxiv.org/abs/2401.00429v1 Deeper and Wider Networks for Performance Metrics Prediction in Communication Networks 2023-12-31T08:59:59Z In today's era, users have increasingly high expectations regarding the performance and efficiency of communication networks. Network operators aspire to achieve efficient network planning, operation, and optimization through Digital Twin Networks (DTN). The effectiveness of DTN heavily relies on the network model, with graph neural networks (GNN) playing a crucial role in network modeling. However, existing network modeling methods still lack a comprehensive understanding of communication networks. In this paper, we propose DWNet (Deeper and Wider Networks), a heterogeneous graph neural network modeling method based on data-driven approaches that aims to address end-to-end latency and jitter prediction in network models. This method stands out due to two distinctive features: firstly, it introduces deeper levels of state participation in the message passing process; secondly, it extensively integrates relevant features during the feature fusion process. Through experimental validation and evaluation, our model achieves higher prediction accuracy compared to previous research achievements, particularly when dealing with unseen network topologies during model training. Our model not only provides more accurate predictions but also demonstrates stronger generalization capabilities across diverse topological structures. 2023-12-31T08:59:59Z Aijia Liu Shiqing Liu Xiaobing Pei http://arxiv.org/abs/2401.00438v1 SFGANS Self-supervised Future Generator for human ActioN Segmentation 2023-12-31T09:36:55Z The ability to locate and classify action segments in long untrimmed video is of particular interest to many applications such as autonomous cars, robotics and healthcare applications. Today, the most popular pipeline for action segmentation is composed of encoding the frames into feature vectors, which are then processed by a temporal model for segmentation. In this paper we present a self-supervised method that comes in the middle of the standard pipeline and generated refined representations of the original feature vectors. Experiments show that this method improves the performance of existing models on different sub-tasks of action segmentation, even without additional hyper parameter tuning. 2023-12-31T09:36:55Z Or Berman Adam Goldbraikh Shlomi Laufer http://arxiv.org/abs/2401.00439v1 On the breathing of spectral bands in periodic quantum waveguides with inflating resonators 2023-12-31T09:38:37Z We are interested in the lower part of the spectrum of the Dirichlet Laplacian $A^\varepsilon$ in a thin waveguide $Π^\varepsilon$ obtained by repeating periodically a pattern, itself constructed by scaling an inner field geometry $Ω$ by a small factor $\varepsilon>0$. The Floquet-Bloch theory ensures that the spectrum of $A^\varepsilon$ has a band-gap structure. Due to the Dirichlet boundary conditions, these bands all move to $+\infty$ as $O(\varepsilon^{-2})$ when $\varepsilon\to0^+$. Concerning their widths, applying techniques of dimension reduction, we show that the results depend on the dimension of the so-called space of almost standing waves in $Ω$ that we denote by $\mathrm{X}_\dagger$. Generically, i.e. for most $Ω$, there holds $\mathrm{X}_\dagger=\{0\}$ and the lower part of the spectrum of $A^\varepsilon$ is very sparse, made of bands of length at most $O(\varepsilon)$ as $\varepsilon\to0^+$. For certain $Ω$ however, we have $\mathrm{dim}\,\mathrm{X}_\dagger=1$ and then there are bands of length $O(1)$ which allow for wave propagation in $Π^\varepsilon$. The main originality of this work lies in the study of the behaviour of the spectral bands when perturbing $Ω$ around a particular $Ω_\star$ where $\mathrm{dim}\,\mathrm{X}_\dagger=1$. We show a breathing phenomenon for the spectrum of $A^\varepsilon$: when inflating $Ω$ around $Ω_\star$, the spectral bands rapidly expand before shrinking. In the process, a band dives below the normalized threshold $π^2/\varepsilon^2$, stops breathing and becomes extremely short as $Ω$ continues to inflate. 2023-12-31T09:38:37Z Lucas Chesnel Sergei A. Nazarov http://arxiv.org/abs/2401.00453v2 Global well-posedness for the Cauchy problem of the Zakharov-Kuznetsov equation on cylindrical spaces 2024-01-02T12:46:37Z We prove that the Zakharov-Kuznetsov equation on cylindrical spaces is globally well-posed below the energy norm. As is known, local well-posedness below energy space was obtained by the first author. We adapt I-method to extend the solutions globally in time. Using modified energies, we obtain the polynomial bounds on the $H^s$ growth for the global solutions. 2023-12-31T11:04:03Z 24 pages Satoshi Osawa Hideo Takaoka http://arxiv.org/abs/2401.00460v1 RainSD: Rain Style Diversification Module for Image Synthesis Enhancement using Feature-Level Style Distribution 2023-12-31T11:30:42Z Autonomous driving technology nowadays targets to level 4 or beyond, but the researchers are faced with some limitations for developing reliable driving algorithms in diverse challenges. To promote the autonomous vehicles to spread widely, it is important to address safety issues on this technology. Among various safety concerns, the sensor blockage problem by severe weather conditions can be one of the most frequent threats for multi-task learning based perception algorithms during autonomous driving. To handle this problem, the importance of the generation of proper datasets is becoming more significant. In this paper, a synthetic road dataset with sensor blockage generated from real road dataset BDD100K is suggested in the format of BDD100K annotation. Rain streaks for each frame were made by an experimentally established equation and translated utilizing the image-to-image translation network based on style transfer. Using this dataset, the degradation of the diverse multi-task networks for autonomous driving, such as lane detection, driving area segmentation, and traffic object detection, has been thoroughly evaluated and analyzed. The tendency of the performance degradation of deep neural network-based perception systems for autonomous vehicle has been analyzed in depth. Finally, we discuss the limitation and the future directions of the deep neural network-based perception algorithms and autonomous driving dataset generation based on image-to-image translation. 2023-12-31T11:30:42Z Under Review Hyeonjae Jeon Junghyun Seo Taesoo Kim Sungho Son Jungki Lee Gyeungho Choi Yongseob Lim http://arxiv.org/abs/2401.00461v1 A Penalized Functional Linear Cox Regression Model for Spatially-defined Environmental Exposure with an Estimated Buffer Distance 2023-12-31T11:31:57Z In environmental health research, it is of interest to understand the effect of the neighborhood environment on health. Researchers have shown a protective association between green space around a person's residential address and depression outcomes. In measuring exposure to green space, distance buffers are often used. However, buffer distances differ across studies. Typically, the buffer distance is determined by researchers a priori. It is unclear how to identify an appropriate buffer distance for exposure assessment. To address geographic uncertainty problem for exposure assessment, we present a domain selection algorithm based on the penalized functional linear Cox regression model. The theoretical properties of our proposed method are studied and simulation studies are conducted to evaluate finite sample performances of our method. The proposed method is illustrated in a study of associations of green space exposure with depression and/or antidepressant use in the Nurses' Health Study. 2023-12-31T11:31:57Z 27 pages, 5 figures Jooyoung Lee Zhibing He Charlotte Roscoe Peter James Li Xu Donna Spiegelman David Zucker Molin Wang http://arxiv.org/abs/2401.00465v1 V2X communication coverage analysis for connected vehicles in intelligent transportation networks: A case study for the city of Xanthi, Greece 2023-12-31T11:39:18Z Intelligent transportation systems (ITS) have been developed to improve traffic flow, efficiency, and safety in transportation. Technological advancements in communication such as the Vehicle-to-Everything (V2X), Vehicle-to-Vehicle (V2V) and Vehicle-to Infrastructure (V2I) enable the real-time exchange of information between vehicles and other entities on the road network, and thus play a significant role in their safety and efficiency. This paper presents a simulation study that models V2V and V2I communication to identify the most suitable range of data transmission between vehicles and infrastructure. The provincial city of Xanthi, Greece is used as a cases study, and the goal is to evaluate whether the proposed placement of Road Side Unit (RSU) provided adequate communication coverage on the city's road network. An analysis through different scenarios identified improvements in traffic management, driving behavior and environmental conditions under different RSU coverage. The results highlight that the communication range of 400 meters is the most adequate option for optimum traffic management in the city of Xanthi. 2023-12-31T11:39:18Z Wireless World Research Forum, Meeting 49, March 28th-30th 2023, Poznań, Poland, Towards sustainable and automated communications Evangelos Bazinas Andreas Gregoriades Marios Raspopoulos Michael Georgiades http://arxiv.org/abs/2401.00478v1 Partial classification of the large-time behavior of solutions to cubic nonlinear Schrödinger systems 2023-12-31T12:44:13Z In this paper, we study the large-time behavior of small solutions to the standard form of the systems of 1D cubic nonlinear Schrödinger equations consisting of two components and possessing a coercive mass-like conserved quantity. The cubic nonlinearity is known to be critical in one space dimension in view of the large-time behavior. By employing the result by Katayama and Sakoda, one can obtain the large-time behavior of the solution if we can integrate the corresponding ODE system. We introduce an integration scheme suited to the system. The key idea is to rewrite the ODE system, which is cubic, as a quadratic system of quadratic quantities of the original unknown. By using this technique, we described the large-time behavior of solutions in terms of elementary functions and the Jacobi elliptic functions for several examples of standard systems. 2023-12-31T12:44:13Z 47 pages, no figure Satoshi Masaki http://arxiv.org/abs/2401.00481v3 Analysis of the $\mathrm{X_{AV}}$ state through its electromagnetic properties 2024-03-26T10:44:18Z To improve our understanding of the quark-gluon dynamics underlying multiquark states, we systematically study their electromagnetic properties. In this study, the magnetic and quadrupole moments of the theoretically predicted singly-charmed state with the quantum numbers $\mathrm{J^P = 1^+}$ is investigated within the framework of the QCD light-cone sum rules method by considering the diquark-antidiquark configuration of this state with quark contents $[ud][\bar{c}\bar{s}]$. The predicted results for the magnetic and quadrupole moments are as $μ_{\mathrm{X_{AV}}}=-0.89 ^{+0.14}_{-0.12}~μ_N $ and $\mathcal{D}_{\mathrm{X_{AV}}} = (-0.46 ^{+0.07}_{-0.06})\times 10^{-2} ~\mbox{fm}^2$. The results obtained can be useful in determining the exact nature of this state. This work will hopefully stimulate experimental interest in the study of the electromagnetic properties of multiquark systems. 2023-12-31T12:46:42Z 8 pages, 1 tables, 1 figure, version accepted by European Physical Journal C U. Özdem http://arxiv.org/abs/2401.00493v1 Reduced variance random batch methods for nonlocal PDEs 2023-12-31T13:27:02Z Random Batch Methods (RBM) for mean-field interacting particle systems enable the reduction of the quadratic computational cost associated with particle interactions to a near-linear cost. The essence of these algorithms lies in the random partitioning of the particle ensemble into smaller batches at each time step. The interaction of each particle within these batches is then evolved until the subsequent time step. This approach effectively decreases the computational cost by an order of magnitude while increasing the amount of fluctuations due to the random partitioning. In this work, we propose a variance reduction technique for RBM applied to nonlocal PDEs of Fokker-Planck type based on a control variate strategy. The core idea is to construct a surrogate model that can be computed on the full set of particles at a linear cost while maintaining enough correlations with the original particle dynamics. Examples from models of collective behavior in opinion spreading and swarming dynamics demonstrate the great potential of the present approach. 2023-12-31T13:27:02Z Lorenzo Pareschi Mattia Zanella http://arxiv.org/abs/2401.00459v3 Interfacial dripping faucet: generating monodisperse liquid lenses 2024-12-17T16:44:10Z We present a surface analog to a dripping faucet, where a viscous liquid slides down an immiscible meniscus. Periodic pinch-off of the dripping filament is observed, generating a succession of monodisperse floating lenses. We show that this interfacial dripping faucet can be described analogously to its single-phase counterpart, replacing surface tension by the spreading coefficient, and even undergoes a transition to a jetting regime. This liquid/liquid/gas system opens perspectives for the study of the dynamics of emulsions at interfaces. 2023-12-31T11:26:13Z 13 pages, 10 figures, 5 supplemental movies Lorène Champougny Vincent Bertin Jacco H. Snoeijer Javier Rodríguez-Rodríguez 10.1103/PhysRevLett.133.254001 http://arxiv.org/abs/2401.00475v3 E-chat: Emotion-sensitive Spoken Dialogue System with Large Language Models 2024-07-27T07:45:43Z This study focuses on emotion-sensitive spoken dialogue in human-machine speech interaction. With the advancement of Large Language Models (LLMs), dialogue systems can handle multimodal data, including audio. Recent models have enhanced the understanding of complex audio signals through the integration of various audio events. However, they are unable to generate appropriate responses based on emotional speech. To address this, we introduce the Emotional chat Model (E-chat), a novel spoken dialogue system capable of comprehending and responding to emotions conveyed from speech. This model leverages an emotion embedding extracted by a speech encoder, combined with LLMs, enabling it to respond according to different emotional contexts. Additionally, we introduce the E-chat200 dataset, designed explicitly for emotion-sensitive spoken dialogue. In various evaluation metrics, E-chat consistently outperforms baseline model, demonstrating its potential in emotional comprehension and human-machine interaction. 2023-12-31T12:29:12Z 5 pages, 3 figures Hongfei Xue Yuhao Liang Bingshen Mu Shiliang Zhang Mengzhe Chen Qian Chen Lei Xie http://arxiv.org/abs/2401.00456v2 Double-well Net for Image Segmentation 2024-07-28T08:40:34Z In this study, our goal is to integrate classical mathematical models with deep neural networks by introducing two novel deep neural network models for image segmentation known as Double-well Nets. Drawing inspirations from the Potts model, our models leverage neural networks to represent a region force functional. We extend the well-know MBO (Merriman-Bence-Osher) scheme to solve the Potts model. The widely recognized Potts model is approximated using a double-well potential and then solved by an operator-splitting method, which turns out to be an extension of the well-known MBO scheme. Subsequently, we replace the region force functional in the Potts model with a UNet-type network, which is data-driven and is designed to capture multiscale features of images, and also introduce control variables to enhance effectiveness. The resulting algorithm is a neural network activated by a function that minimizes the double-well potential. What sets our proposed Double-well Nets apart from many existing deep learning methods for image segmentation is their strong mathematical foundation. They are derived from the network approximation theory and employ the MBO scheme to approximately solve the Potts model. By incorporating mathematical principles, Double-well Nets bridge the MBO scheme and neural networks, and offer an alternative perspective for designing networks with mathematical backgrounds. Through comprehensive experiments, we demonstrate the performance of Double-well Nets, showcasing their superior accuracy and robustness compared to state-of-the-art neural networks. Overall, our work represents a valuable contribution to the field of image segmentation by combining the strengths of classical variational models and deep neural networks. The Double-well Nets introduce an innovative approach that leverages mathematical foundations to enhance segmentation performance. 2023-12-31T11:16:12Z Hao Liu Jun Liu Raymond H. Chan Xue-Cheng Tai http://arxiv.org/abs/2401.00451v1 Exploring the Need of Accessibility Education in the Software Industry: Insights from a Survey of Software Professionals in India 2023-12-31T10:58:30Z A UserWay study in 2021 indicates that an annual global e-commerce revenue loss of approximately $16 billion can be attributed to inaccessible websites and applications. According to the 2023 WebAIM study, only 3.7% of the world's top one million website homepages are fully accessible. This shows that many software developers use poor coding practices that don't adhere to the Web Content Accessibility Guidelines (WCAG). This research centers on software professionals and their role in addressing accessibility. This work seeks to understand (a) who within the software development community actively practices accessibility, (b) when and how accessibility is considered in the software development lifecycle, (c) the various challenges encountered in building accessible software, and (d) the resources required by software professionals to enhance product accessibility. Our survey of 269 software professionals from India sheds light on the pressing need for accessibility education within the software industry. A substantial majority (69.9%, N=269) of respondents express the need for training materials, workshops, and bootcamps to enhance their accessibility skills. We present a list of actionable recommendations that can be implemented within the industry to promote accessibility awareness and skills. We also open source our raw data for further research, encouraging continued exploration in this domain. 2023-12-31T10:58:30Z To be published in International Conference on Software Engineering (ICSE'24), Software Engineering Education and Training Track Parthasarathy P D Swaroop Joshi 10.1145/3639474.3640079 http://arxiv.org/abs/2401.00433v4 Quantized collision invariants on the sphere 2024-04-20T18:29:05Z We show that a measurable function $g:\mathbb{S}^{d-1}\to\mathbb{R}$, with $d\geq 3$, satisfies the functional relation \begin{equation*} g(ω)+g(ω_*)=g(ω')+g(ω_*'), \end{equation*} for all admissible $ω,ω_*,ω',ω_*'\in\mathbb{S}^{d-1}$ in the sense that \begin{equation*} ω+ω_*=ω'+ω_*', \end{equation*} if and only if it can be written as \begin{equation*} g(ω)=A+B\cdotω, \end{equation*} for some constants $A\in \mathbb{R}$ and $B\in\mathbb{R}^d$. Such functions form a family of quantized collision invariants which play a fundamental role in the study of hydrodynamic regimes of the Boltzmann--Fermi--Dirac equation near Fermionic condensates, i.e., at low temperatures. In particular, they characterize the elastic collisional dynamics of Fermions near a statistical equilibrium where quantum effects are predominant. 2023-12-31T09:15:00Z Communications in Mathematics, Volume 32 (2024), Issue 3 (Special issue: Portuguese Mathematics) (April 25, 2024) cm:12766 Benjamin Anwasia Diogo Arsénio 10.46298/cm.12766 http://arxiv.org/abs/2401.00497v2 The vector-valued Stieltjes moment problem with general exponents 2024-06-24T07:43:37Z We characterize the sequences of complex numbers $(z_{n})_{n \in \mathbb{N}}$ and the locally complete $(DF)$-spaces $E$ such that for each $(e_{n})_{n \in \mathbb{N}} \in E^\mathbb{N}$ there exists an $E$-valued function $\mathbf{f}$ on $(0,\infty)$ (satisfying a mild regularity condition) such that $$\int_{0}^{\infty} t^{z_{n}} \mathbf{f}(t) dt = e_{n}, \qquad \forall n \in \mathbb{N},$$ where the integral should be understood as a Pettis integral. Moreover, in this case, we show that there always exists a solution $\mathbf{f}$ that is smooth on $(0,\infty)$ and satisfies certain optimal growth bounds near $0$ and $\infty$. The scalar-valued case $(E = \mathbb{C})$ was treated by Durán [Math. Nachr. 158 (1992), 175-194]. Our work is based upon his result. 2023-12-31T13:35:11Z 13 pages Andreas Debrouwere Lenny Neyt 10.1007/s43037-024-00364-8 http://arxiv.org/abs/2401.00448v3 Beyond Chinchilla-Optimal: Accounting for Inference in Language Model Scaling Laws 2025-04-14T10:11:13Z Large language model (LLM) scaling laws are empirical formulas that estimate changes in model quality as a result of increasing parameter count and training data. However, these formulas, including the popular Deepmind Chinchilla scaling laws, neglect to include the cost of inference. We modify the Chinchilla scaling laws to calculate the optimal LLM parameter count and pre-training data size to train and deploy a model of a given quality and inference demand. We conduct our analysis both in terms of a compute budget and real-world costs and find that LLM researchers expecting reasonably large inference demand (~1B requests) should train models smaller and longer than Chinchilla-optimal. Furthermore, we train 47 models of varying sizes and parameter counts to validate our formula and find that model quality continues to improve as we scale tokens per parameter to extreme ranges (up to 10,000). Finally, we ablate the procedure used to fit the Chinchilla scaling law coefficients and find that developing scaling laws only from data collected at typical token/parameter ratios overestimates the impact of additional tokens at these extreme ranges. 2023-12-31T10:53:58Z 16 pages, 7 figures, In the 41st International Conference on Machine Learning, 2024 Nikhil Sardana Jacob Portes Sasha Doubov Jonathan Frankle http://arxiv.org/abs/2401.00431v2 Wild2Avatar: Rendering Humans Behind Occlusions 2025-08-14T22:41:09Z Rendering the visual appearance of moving humans from occluded monocular videos is a challenging task. Most existing research renders 3D humans under ideal conditions, requiring a clear and unobstructed scene. Those methods cannot be used to render humans in real-world scenes where obstacles may block the camera's view and lead to partial occlusions. In this work, we present Wild2Avatar, a neural rendering approach catered for occluded in-the-wild monocular videos. We propose occlusion-aware scene parameterization for decoupling the scene into three parts - occlusion, human, and background. Additionally, extensive objective functions are designed to help enforce the decoupling of the human from both the occlusion and the background and to ensure the completeness of the human model. We verify the effectiveness of our approach with experiments on in-the-wild videos. 2023-12-31T09:01:34Z IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI). Webpage: https://cs.stanford.edu/~xtiange/projects/wild2avatar/ Tiange Xiang Adam Sun Scott Delp Kazuki Kozuka Li Fei-Fei Ehsan Adeli http://arxiv.org/abs/2401.00422v3 Interpreting the Curse of Dimensionality from Distance Concentration and Manifold Effect 2025-03-20T10:08:31Z The characteristics of data like distribution and heterogeneity, become more complex and counterintuitive as dimensionality increases. This phenomenon is known as curse of dimensionality, where common patterns and relationships (e.g., internal pattern and boundary pattern) that hold in low-dimensional space may be invalid in higher-dimensional space. It leads to a decreasing performance for the regression, classification, or clustering models or algorithms. Curse of dimensionality can be attributed to many causes. In this paper, we first summarize the potential challenges associated with manipulating high-dimensional data, and explains the possible causes for the failure of regression, classification, or clustering tasks. Subsequently, we delve into two major causes of the curse of dimensionality, distance concentration, and manifold effect, by performing theoretical and empirical analyses. The results demonstrate that, as the dimensionality increases, nearest neighbor search (NNS) using three classical distance measurements, Minkowski distance, Chebyshev distance, and cosine distance, becomes meaningless. Meanwhile, the data incorporates more redundant features, and the variance contribution of principal component analysis (PCA) is skewed towards a few dimensions. 2023-12-31T08:22:51Z 21 pages, 10 figures Dehua Peng Zhipeng Gui Huayi Wu http://arxiv.org/abs/2401.00410v1 Electrical and thermal transport properties of kagome metals AV$_3$Sb$_5$ (A=K, Rb, Cs) 2023-12-31T06:48:14Z The interplay between lattice geometry, band topology and electronic correlations in the newly discovered kagome compounds AV$_3$Sb$_5$ (A=K, Rb, Cs) makes this family a novel playground to investigate emergent quantum phenomena, such as unconventional superconductivity, chiral charge density wave and electronic nematicity. These exotic quantum phases naturally leave nontrivial fingerprints in transport properties of AV$_3$Sb$_5$, both in electrical and thermal channels, which are prominent probes to uncover the underlying mechanisms. In this brief review, we highlight the unusual electrical and thermal transport properties observed in the unconventional charge ordered state of AV3Sb5, including giant anomalous Hall, anomalous Nernst, ambipolar Nernst and anomalous thermal Hall effects. Connections of these anomalous transport properties to time-reversal symmetry breaking, topological and multiband fermiology, as well as electronic nematicity, are also discussed. Finally, a perspective together with challenges of this rapid growing field are given. 2023-12-31T06:48:14Z 34 pages,9 figures,an review article published in Tungsten 5,300(2023) Tungsten 5,300(2023) Xinrun Mi Kunya Yang Yuhan Gan Long Zhang Aifeng Wang Yisheng Chai Xiaoyuan Zhou Mingquan He 10.1007/s42864-022-00192-z http://arxiv.org/abs/2401.00492v2 Edge statistics for random band matrices 2025-06-03T07:35:20Z We consider Hermitian and symmetric random band matrices on the $d$-dimensional lattice $(\mathbb{Z}/L\mathbb{Z})^d$ with bandwidth $W$, focusing on local eigenvalue statistics at the spectral edge in the limit $W\to\infty$. Our analysis reveals a critical dimension $d_c=6$ and identifies the critical bandwidth scaling as $W_c=L^{(1-d/6)_+}$. In the Hermitian case, we establish the Anderson transition for all dimensions $d<4$, and GUE edge universality when $d\geq 4$ under the condition $W\geq L^{1/3+ε}$ for any $ε>0$. In the symmetric case, we also establish parallel but more subtle transition phenomena after tadpole diagram renormalization. These findings extend Sodin's pioneering work [Ann. Math. 172, 2010], which was limited to the one-dimensional case and did not address the critical phenomena. 2023-12-31T13:25:37Z Page 88, figures 14, Added Section 5 on tadpole diagram renormalization enables our results as better for symmetric case to as for Hermitian; We remove power-law band matrices that will be treated in a separate paper Dang-Zheng Liu Guangyi Zou http://arxiv.org/abs/2401.00415v1 A Novel Estimation Method for Temperature of Magnetic Nanoparticles Dominated by Brownian Relaxation Based on Magnetic Particle Spectroscopy 2023-12-31T07:29:31Z This paper presents a novel method for estimating the temperature of magnetic nanoparticles (MNPs) based on AC magnetization harmonics of MNPs dominated by Brownian relaxation. The difference in the AC magnetization response and magnetization harmonic between the Fokker-Planck equation and the Langevin function was analyzed, and we studied the relationship between the magnetization harmonic and the key factors, such as Brownian relaxation time, temperature, magnetic field strength, core size and hydrodynamic size of MNPs, excitation frequency, and so on. We proposed a compensation function for AC magnetization harmonic with consideration of the key factors and the difference between the Fokker-Planck equation and the Langevin function. Then a temperature estimation model based on the compensation function and the Langevin function was established. By employing the least squares algorithm, the temperature was successfully calculated. The experimental results show that the temperature error is less than 0.035 K in the temperature range from 310 K to 320 K. The temperature estimation model is expected to improve the performance of the magnetic nanoparticle thermometer and be applied to magnetic nanoparticle-mediated hyperthermia. 2023-12-31T07:29:31Z Zhongzhou Du Gaoli Zhao Zhanpeng Hua Na Ye Yi Sun Wenjie Wu Haochen Zhang Longtu Yu Shijie Han Haozhe Wang Wenzhong Liu Takashi Yoshida http://arxiv.org/abs/2401.00419v1 Approximation algorithms for Job Scheduling with reconfigurable resources 2023-12-31T07:50:58Z We consider here the MultiBot problem for the scheduling and the resource parametrization of jobs related to the production or the transportation of different products inside a given time horizon. Those jobs must meet known in advance demands. The time horizon is divided into several discrete identical periods representing each the time needed to proceed a job. The objective is to find a parametrization and a schedule for the jobs in such a way they require as less resources as possible. Though this problem derived from the applicative context of reconfigurable robots, we focus here on fundamental issues. We show that the resulting strongly NP-hard Multibot problem may be handled in a greedy way with an approximation ratio of $\frac{4}{3}$. 2023-12-31T07:50:58Z 26 pages, 4 figures Pierre Bergé Mari Chaikovskaia Jean-Philippe Gayon Alain Quilliot http://arxiv.org/abs/2401.00430v2 Brain-Conditional Multimodal Synthesis: A Survey and Taxonomy 2024-01-03T08:50:27Z In the era of Artificial Intelligence Generated Content (AIGC), conditional multimodal synthesis technologies (e.g., text-to-image, text-to-video, text-to-audio, etc) are gradually reshaping the natural content in the real world. The key to multimodal synthesis technology is to establish the mapping relationship between different modalities. Brain signals, serving as potential reflections of how the brain interprets external information, exhibit a distinctive One-to-Many correspondence with various external modalities. This correspondence makes brain signals emerge as a promising guiding condition for multimodal content synthesis. Brian-conditional multimodal synthesis refers to decoding brain signals back to perceptual experience, which is crucial for developing practical brain-computer interface systems and unraveling complex mechanisms underlying how the brain perceives and comprehends external stimuli. This survey comprehensively examines the emerging field of AIGC-based Brain-conditional Multimodal Synthesis, termed AIGC-Brain, to delineate the current landscape and future directions. To begin, related brain neuroimaging datasets, functional brain regions, and mainstream generative models are introduced as the foundation of AIGC-Brain decoding and analysis. Next, we provide a comprehensive taxonomy for AIGC-Brain decoding models and present task-specific representative work and detailed implementation strategies to facilitate comparison and in-depth analysis. Quality assessments are then introduced for both qualitative and quantitative evaluation. Finally, this survey explores insights gained, providing current challenges and outlining prospects of AIGC-Brain. Being the inaugural survey in this domain, this paper paves the way for the progress of AIGC-Brain research, offering a foundational overview to guide future work. 2023-12-31T09:00:40Z Weijian Mai Jian Zhang Pengfei Fang Zhijun Zhang http://arxiv.org/abs/2401.00432v1 Magnetic order and strongly-correlated effects in the one-dimensional Ising-Kondo lattice 2023-12-31T09:04:36Z We investigate the magnetic order and related strongly-correlated effects in an one-dimensional Ising-Kondo lattice with transverse field. This model is the anisotropic limit of the conventional isotropic Kondo lattice model, in the sense that the itinerant electrons interact with the localized magnetic moments via only longitudinal Kondo exchange. Adopting the numerical density-matrix-renormalization group method, we map out the ground-state phase diagram in various parameter spaces. Depending on the Kondo coupling and filling number, three distinct phases, including a metallic paramagnetic, a metallic ferromagnetic, and a gapped spin-density wave phase, are obtained. The spin-density wave is characterized by an ordering wave vector which coincides with the nesting wave vector of the Fermi surface. This makes the corresponding magnetic transition a spin analog of the Peierls transition occurring in the one-dimensional metal. Moreover, by analyzing the momentum distribution function and charge correlation function, the conduction electrons are shown to behave like free spinless fermions in the ferromagnetic phase. We finally discuss the effect of the repulsive Hubbard interaction between conduction electrons. Our work enriches the Kondo physics and deepens the current understanding of the heavy fermion compounds. 2023-12-31T09:04:36Z 9 pages, 12 figures Xiaofan Zhou Jingtao Fan Suotang Jia http://arxiv.org/abs/2401.00434v2 GeoGalactica: A Scientific Large Language Model in Geoscience 2024-04-13T17:05:03Z Large language models (LLMs) have achieved huge success for their general knowledge and ability to solve a wide spectrum of tasks in natural language processing (NLP). Due to their impressive abilities, LLMs have shed light on potential inter-discipline applications to foster scientific discoveries of a specific domain by using artificial intelligence (AI for science, AI4S). In the meantime, utilizing NLP techniques in geoscience research and practice is wide and convoluted, contributing from knowledge extraction and document classification to question answering and knowledge discovery. In this work, we take the initial step to leverage LLM for science, through a rather straightforward approach. We try to specialize an LLM into geoscience, by further pre-training the model with a vast amount of texts in geoscience, as well as supervised fine-tuning (SFT) the resulting model with our custom collected instruction tuning dataset. These efforts result in a model GeoGalactica consisting of 30 billion parameters. To our best knowledge, it is the largest language model for the geoscience domain. More specifically, GeoGalactica is from further pre-training of Galactica. We train GeoGalactica over a geoscience-related text corpus containing 65 billion tokens, preserving as the largest geoscience-specific text corpus. Then we fine-tune the model with 1 million pairs of instruction-tuning data consisting of questions that demand professional geoscience knowledge to answer. In this technical report, we will illustrate in detail all aspects of GeoGalactica, including data collection, data cleaning, base model selection, pre-training, SFT, and evaluation. We open-source our data curation tools and the checkpoints of GeoGalactica during the first 3/4 of pre-training. 2023-12-31T09:22:54Z Zhouhan Lin Cheng Deng Le Zhou Tianhang Zhang Yi Xu Yutong Xu Zhongmou He Yuanyuan Shi Beiya Dai Yunchong Song Boyi Zeng Qiyuan Chen Yuxun Miao Bo Xue Shu Wang Luoyi Fu Weinan Zhang Junxian He Yunqiang Zhu Xinbing Wang Chenghu Zhou http://arxiv.org/abs/2401.00445v1 Energy-Efficient Power Control for Multiple-Task Split Inference in UAVs: A Tiny Learning-Based Approach 2023-12-31T10:16:59Z The limited energy and computing resources of unmanned aerial vehicles (UAVs) hinder the application of aerial artificial intelligence. The utilization of split inference in UAVs garners significant attention due to its effectiveness in mitigating computing and energy requirements. However, achieving energy-efficient split inference in UAVs remains complex considering of various crucial parameters such as energy level and delay constraints, especially involving multiple tasks. In this paper, we present a two-timescale approach for energy minimization in split inference, where discrete and continuous variables are segregated into two timescales to reduce the size of action space and computational complexity. This segregation enables the utilization of tiny reinforcement learning (TRL) for selecting discrete transmission modes for sequential tasks. Moreover, optimization programming (OP) is embedded between TRL's output and reward function to optimize the continuous transmit power. Specifically, we replace the optimization of transmit power with that of transmission time to decrease the computational complexity of OP since we reveal that energy consumption monotonically decreases with increasing transmission time. The replacement significantly reduces the feasible region and enables a fast solution according to the closed-form expression for optimal transmit power. Simulation results show that the proposed algorithm can achieve a higher probability of successful task completion with lower energy consumption. 2023-12-31T10:16:59Z Chenxi Zhao Min Sheng Junyu Liu Tianshu Chu Jiandong Li http://arxiv.org/abs/2401.00446v1 Dissipation of AGN jets in a clumpy interstellar medium 2023-12-31T10:29:38Z Accreting supermassive black holes (SMBHs) frequently power jets that interact with the interstellar/circumgalactic medium (ISM/CGM), regulating star-formation in the galaxy. Highly supersonic jets launched by active galactic nuclei (AGN) power a cocoon that confines them and shocks the ambient medium. We build upon the models of narrow conical jets interacting with a smooth ambient medium, to include the effect of dense clouds that are an essential ingredient of a multiphase ISM. The key physical ingredient of this model is that the clouds along the supersonic jet-beam strongly decelerate the jet-head, but the subsonic cocoon easily moves around the clouds without much resistance. We propose scalings for important physical quantities -- cocoon pressure, head & cocoon speed, and jet radius. We obtain, for the first time, the analytic condition on clumpiness of the ambient medium for the jet to dissipate within the cocoon and verify it with numerical simulations of conical jets interacting with a uniform ISM with embedded spherical clouds. A jet is defined to be dissipated when the cocoon speed exceeds the speed of the jet-head. We compare our models to more sophisticated numerical simulations, direct observations of jet-ISM interaction (e.g., quasar J1316+1753), and discuss implications for the Fermi/eROSITA bubbles. Our work also motivates effective subgrid models for AGN jet feedback in a clumpy ISM unresolved by the present generation of cosmological galaxy formation simulations. 2023-12-31T10:29:38Z 23 pages, 12 figures, 3 tables; to be submitted; comments are welcome; accompanying video: http://youtu.be/DUpSwMMrGfk Riju Dutta Prateek Sharma Kartick C. Sarkar James M. Stone http://arxiv.org/abs/2401.00470v2 More on $G$-flux and General Hodge Cycles on the Fermat Sextic 2024-02-26T13:32:24Z We study M-Theory solutions with $G$-flux on the Fermat sextic Calabi-Yau fourfold, focussing on the relationship between the number of stabilized complex structure moduli and the tadpole contribution of the flux. We use two alternative approaches to define the fluxes: algebraic cycles and (appropriately quantized) Griffiths residues. In both cases, we collect evidence for the non-existence of solutions which stabilize all moduli and stay within the tadpole bound 2023-12-31T11:58:14Z v2: typos corrected and references added Andreas P. Braun Hugo Fortin Daniel Lopez Garcia Roberto Villaflor Loyola http://arxiv.org/abs/2401.00472v1 On Thurston's geometrical space form problem: on quasi space forms 2023-12-31T12:03:20Z A proposal is made for what may well be the most elementary Riemannian spaces which are homogeneous but not isotropic. In other words: a proposal is made for what may well be the nicest symmetric spaces beyond the real space forms, that is, beyond the Riemannian spaces which are homogeneous and isotropic. The above qualification of `'nicest symmetric spaces'' finds a justification in that, together with the real space forms, these spaces are most natural with respect to the importance in human vision of our ability to readily recognise conformal things and in that these spaces are most natural with respect to what in Weyl's view is symmetry in Riemannian geometry. Following his suggestion to remove the real space forms' isotropy condition, the quasi space forms thus introduced do offer a metrical, local geometrical solution to the geometrical space form problem as posed by Thurston in his 1979 Princeton Lecture Notes on `'The Geometry and Topology of 3-manifolds''. Roughly speaking, quasi space forms are the Riemannian manifolds of dimension greater than or equal to 3, which are not real space forms but which admit two orthogonally complementary distributions such that at all points all the 2-planes that in the tangent spaces there are situated in a same position relative to these distributions do have the same sectional curvatures. 2023-12-31T12:03:20Z 20 pages Stefan Haesen Miroslava Petrović-Torgašev Leopold Verstraelen http://arxiv.org/abs/2401.00482v1 Structural deformation and irreversible magnetic properties of flexible Co/Pt and Co/Pd thin films 2023-12-31T12:55:59Z The successful commercialization of flexible spintronic devices requires a complete understanding of the impact of external strain on the structural, electronic, and magnetic properties of a system. The impact of bending-induced strain on flexible films is studied quite well. However, little is known about the effect of other modes of flexibility, e.g., wrinkling, twisting, peeling, and stretching on the functional properties of flexible films. In this context, perpendicular magnetic anisotropic Co/Pt and Co/Pd thin films are prepared on flexible Kapton substrates, and the impact of the peeling mode is studied in detail. The peeling method generates numerous cracks, and buckling in the thin film, along with localized blister formation imaged by scanning electron microscopy. Further, the resistivity measurement confirms a significant enhancement in sample resistance owing to the severe damage of the films. The structural discontinuities strongly affect the magnetization reversal phenomena as measured by the magneto-optic Kerr effect (MOKE)-based microscopy. The bubble domains got converted to elongated-shaped domains due to several hindrances to the wall motion after strain application. Further, the relaxation measurements reveal that the thermal energy is insufficient to switch the magnetization at a few areas due to their high pinning potential associated with the damages. In contrast to bending-induced strain, here, all the modifications in the functional properties are found to be irreversible in nature. 2023-12-31T12:55:59Z Esita Pandey Shaktiranjan Mohanty Abhisek Mishra Bhuvneshwari Sharma Subhankar Bedanta http://arxiv.org/abs/2401.00484v1 Directional flow in perivascular networks: Mixed finite elements for reduced-dimensional models on graphs 2023-12-31T12:57:35Z The flow of cerebrospinal fluid through the perivascular spaces of the brain is believed to play a crucial role in eliminating toxic waste proteins. While the driving forces of this flow have been enigmatic, experiments have shown that arterial wall motion is central. In this work, we present a network model for simulating pulsatile fluid flow in perivascular networks. We establish the well-posedness of this model in the primal and dual mixed variational settings, and show how it can be discretized using mixed finite elements. Further, we utilize this model to investigate fundamental questions concerning the physical mechanisms governing perivascular fluid flow. Notably, our findings reveal that arterial pulsations can induce directional flow in branching perivascular networks. 2023-12-31T12:57:35Z Ingeborg G. Gjerde Miroslav Kuchta Marie E. Rognes Barbara Wohlmuth http://arxiv.org/abs/2401.00486v1 Molecular Hybridization Induced Antidamping and Sizable Enhanced Spin-to-Charge Conversion in Co20Fe60B20/$β$-W/C60 Heterostructures 2023-12-31T12:59:23Z Development of power efficient spintronics devices has been the compelling need in the post-CMOS technology era. The effective tunability of spin-orbit-coupling (SOC) in bulk and at the interfaces of hybrid materials stacking is a prerequisite for scaling down the dimension and power consumption of these devices. In this work, we demonstrate the strong chemisorption of C60 molecules when grown on the high SOC $β$-W layer. The parent CFB/$β$-W bilayer exhibits large spin-to-charge interconversion efficiency, which can be ascribed to the interfacial SOC observed at the Ferromagnet/Heavy metal interface. Further, the adsorption of C60 molecules on $β$-W reduces the effective Gilbert damping by $\sim$15% in the CFB/$β$-W/C60 heterostructures. The anti-damping is accompanied by a gigantic $\sim$115% enhancement in the spin-pumping induced output voltage owing to the molecular hybridization. The non-collinear Density Functional Theory calculations confirm the long-range enhancement of SOC of $β$-W upon the chemisorption of C60 molecules, which in turn can also enhance the SOC at the CFB/$β$-W interface in CFB/$β$-W/C60 heterostructures. The combined amplification of bulk as well interfacial SOC upon molecular hybridization stabilizes the anti-damping and enhanced spin-to-charge conversion, which can pave the way for the fabrication of power efficient spintronics devices. 2023-12-31T12:59:23Z Antarjami Sahoo Aritra Mukhopadhyaya Swayang Priya Mahanta Md. Ehesan Ali Subhankar Bedanta http://arxiv.org/abs/2401.00491v1 On the sense of convergence in the dyadic representation theorem 2023-12-31T13:21:37Z The dyadic representation of any singular integral operator, as an average of dyadic model operators, has found many applications. While for many purposes it is enough to have such a representation for a "suitable class" of test functions, we show that, under quite general assumptions (essentially minimal ones to make sense of the formula), the representation is actually valid for all pairs $(f,g)\in L^p(\mathbb R^d)\times L^{p'}(\mathbb R^d)$, not just test functions. 2023-12-31T13:21:37Z 25 pages Tuomas Hytönen http://arxiv.org/abs/2401.00496v2 SAR-RARP50: Segmentation of surgical instrumentation and Action Recognition on Robot-Assisted Radical Prostatectomy Challenge 2024-01-23T23:30:57Z Surgical tool segmentation and action recognition are fundamental building blocks in many computer-assisted intervention applications, ranging from surgical skills assessment to decision support systems. Nowadays, learning-based action recognition and segmentation approaches outperform classical methods, relying, however, on large, annotated datasets. Furthermore, action recognition and tool segmentation algorithms are often trained and make predictions in isolation from each other, without exploiting potential cross-task relationships. With the EndoVis 2022 SAR-RARP50 challenge, we release the first multimodal, publicly available, in-vivo, dataset for surgical action recognition and semantic instrumentation segmentation, containing 50 suturing video segments of Robotic Assisted Radical Prostatectomy (RARP). The aim of the challenge is twofold. First, to enable researchers to leverage the scale of the provided dataset and develop robust and highly accurate single-task action recognition and tool segmentation approaches in the surgical domain. Second, to further explore the potential of multitask-based learning approaches and determine their comparative advantage against their single-task counterparts. A total of 12 teams participated in the challenge, contributing 7 action recognition methods, 9 instrument segmentation techniques, and 4 multitask approaches that integrated both action recognition and instrument segmentation. The complete SAR-RARP50 dataset is available at: https://rdr.ucl.ac.uk/projects/SARRARP50_Segmentation_of_surgical_instrumentation_and_Action_Recognition_on_Robot-Assisted_Radical_Prostatectomy_Challenge/191091 2023-12-31T13:32:18Z Dimitrios Psychogyios Emanuele Colleoni Beatrice Van Amsterdam Chih-Yang Li Shu-Yu Huang Yuchong Li Fucang Jia Baosheng Zou Guotai Wang Yang Liu Maxence Boels Jiayu Huo Rachel Sparks Prokar Dasgupta Alejandro Granados Sebastien Ourselin Mengya Xu An Wang Yanan Wu Long Bai Hongliang Ren Atsushi Yamada Yuriko Harai Yuto Ishikawa Kazuyuki Hayashi Jente Simoens Pieter DeBacker Francesco Cisternino Gabriele Furnari Alex Mottrie Federica Ferraguti Satoshi Kondo Satoshi Kasai Kousuke Hirasawa Soohee Kim Seung Hyun Lee Kyu Eun Lee Hyoun-Joong Kong Kui Fu Chao Li Shan An Stefanie Krell Sebastian Bodenstedt Nicolas Ayobi Alejandra Perez Santiago Rodriguez Juanita Puentes Pablo Arbelaez Omid Mohareri Danail Stoyanov http://arxiv.org/abs/2401.00401v1 Multiplayer Battle Game-Inspired Optimizer for Complex Optimization Problems 2023-12-31T05:28:12Z Various popular multiplayer battle royale games share a lot of common elements. Drawing from our observations, we summarized these shared characteristics and subsequently proposed a novel heuristic algorithm named multiplayer battle game-inspired optimizer (MBGO). The proposed MBGO streamlines mainstream multiplayer battle royale games into two discrete phases: movement and battle. Specifically, the movement phase incorporates the principles of commonly encountered ``safe zones'' to incentivize participants to relocate to areas with a higher survival potential. The battle phase simulates a range of strategies adopted by players in various situations to enhance the diversity of the population. To evaluate and analyze the performance of the proposed MBGO, we executed it alongside eight other algorithms, including three classics and five latest ones, across multiple diverse dimensions within the CEC2017 and CEC2020 benchmark functions. In addition, we employed several industrial design problems to evaluate the scalability and practicality of the proposed MBGO. The results of the statistical analysis reveal that the novel MBGO demonstrates significant competitiveness, excelling not only in convergence speed, but also in achieving high levels of convergence accuracy across both benchmark functions and real-world problems. 2023-12-31T05:28:12Z Yuefeng Xu Rui Zhong Chao Zhang Jun Yu http://arxiv.org/abs/2401.00489v2 Hodge Theory of Abelian Covers 2024-07-17T14:45:23Z Motivated by classical Alexander invariants of affine hypersurface complements, we endow certain finite dimensional quotients of the homology of abelian covers of complex algebraic varieties with a canonical and functorial mixed Hodge structure (MHS). More precisely, we focus on covers which arise algebraically in the following way: if $U$ is a smooth connected complex algebraic variety and $G$ is a complex semiabelian variety, the pullback of the exponential map by an algebraic morphism $f:U\to G$ yields a covering space $π:U^f\to U$ whose group of deck transformations is $π_1(G)$. The new MHSs are compatible with Deligne's MHS on the homology of $U$ through the covering map $π$ and satisfy a direct sum decomposition as MHSs into generalized eigenspaces by the action of deck transformations. This provides a vast generalization of the previous results regarding univariable Alexander modules by Geske, Maxim, Wang and the authors. Lastly, we reduce the problem of whether the first Betti number of the Milnor fiber of a central hyperplane arrangement complement is combinatorial to a question about the Hodge filtration of certain MHSs defined in this paper, providing evidence that the new structures contain interesting information. 2023-12-31T13:18:06Z 72 pages Eva Elduque Moisés Herradón Cueto http://arxiv.org/abs/2401.00407v2 CMB lensing from early-formed dark matter halos 2024-05-20T01:11:30Z Some theoretical models for the early Universe predict a spike-type enhancement in the primordial power spectrum on a small scale, which would result in forming early-formed dark matter halos~(EFHs). Some recent studies have claimed to have placed limits on such small scales, which, however, involve uncertainties, such as the physics of substructures and the halo-galaxy relations. In this work, we study the cosmic microwave background~(CMB) lensing effect, considering the existence of EFHs, and investigate the potential to probe the EFHs and the primordial perturbations on scales smaller than $1\mathrm{Mpc}$, complementing these previous studies. We numerically calculate the angular power spectrum of the lensing potential and the lensed CMB anisotropy of temperature, E-mode, and B-mode polarization, including the nonlinear effects of EFHs. We find the possibility that the lensed CMB temperature anisotropy is significantly enhanced on small scales, $\ell>1000$, and could be tested by component decomposition of observed signals through multifrequency observations. Through the calculation with different models of the spiky-type power spectrum, we demonstrate that the accurate measurements of the CMB lensing effect would provide insight into the abundance of EFHs within the limited mass range around $10^{11}~M_\odot$ and the primordial power spectrum on the limited scales around $k\sim 1\mathrm{Mpc}^{-1}$. In particular, we find that the existence of such EFHs can amplify the lensed anisotropy of CMB B-mode polarization even on large scales, $\ell <100$, as the overall enhancement by $\sim 5 \%$ level compared to the standard structure formation model without EFHs. Therefore, future CMB measurements, such as the LiteBIRD satellite, can probe the existence of the EFHs and the spike-type primordial power spectrum through the precise measurement of the large-scale CMB B-mode polarization. 2023-12-31T05:47:18Z 14 pages, 7 figures Phys. Rev. D 109, (2024) 103524 Katsuya T. Abe Hiroyuki Tashiro 10.1103/PhysRevD.109.103524 http://arxiv.org/abs/2401.00464v3 A note on the Lp-Sobolev inequality 2024-11-11T13:49:10Z The usual Sobolev inequality in $\mathbb{R}^N$, asserts that $\|\nabla u\|_{L^p(\mathbb{R}^N)} \geq \mathcal{S}\|u\|_{L^{p^*}(\mathbb{R}^N)}$ for $1<p<N$ and $p^*=\frac{pN}{N-p}$, with $\mathcal{S}$ being the sharp constant. Based on a recent work of Figalli and Zhang [Duke Math. J., 2022], a weak norm remainder term of Sobolev inequality in a subdomain $Ω\subset \mathbb{R}^N$ with finite measure is established, i.e., for $\frac{2N}{N+1}<p<N$ there exists a constant $\mathcal{C}>0$ independent of $Ω$ such that \[ \|\nabla u\|^p_{L^p(Ω)} -\mathcal{S}^p\|u\|^p_{L^{p^*}(Ω)} \geq \mathcal{C}|Ω|^{-\fracγ{p^*(p-1)}} \|u\|_{L^{\bar{p}}_w(Ω)}^γ\| u\|_{L^{p^*}(Ω)}^{p-γ},\quad \mbox{for all}\ u\in C^\infty_0(Ω)\setminus\{0\}, \] where $γ=\max\{2,p\}$, $\bar{p}=p^*(p-1)/p$, and $\|\cdot\|_{L^{\bar{p}}_w(Ω)}$ denotes the weak $L^{\bar{p}}$-norm. Moreover, we establish a sharp upper bound of Sobolev inequality in $\mathbb{R}^N$. 2023-12-31T11:39:03Z Shengbing Deng Xingliang Tian http://arxiv.org/abs/2401.00455v2 Worldtube puncture scheme for first- and second-order self-force calculations in the Fourier domain 2024-10-26T18:56:05Z Second-order gravitational self-force theory has recently led to the breakthrough calculation of ``first post-adiabatic'' (1PA) compact-binary waveforms [Phys. Rev. Lett. 130, 241402 (2023)]. The computations underlying those waveforms depend on a method of solving the perturbative second-order Einstein equation on a Schwarzschild background in the Fourier domain. In this paper we present that method, which involves dividing the domain into several regions. Different regions utilize different time slicings and allow for the use of ``punctures'' to tame sources and enforce physical boundary conditions. We demonstrate the method for Lorenz-gauge and Teukolsky equations in the relatively simple case of calculating parametric derivatives (``slow time derivatives'') of first-order fields, which are an essential input at second order. 2023-12-31T11:15:41Z 41 pages, 10 figures. Minor changes in response to referee Phys. Rev. D 109, 104010 (2024) Jeremy Miller Benjamin Leather Adam Pound Niels Warburton http://arxiv.org/abs/2401.00480v3 Modelling contagious viral dynamics: a kinetic approach based on mutual utility 2024-02-09T11:04:16Z The temporal evolution of a contagious viral disease is modelled as the dynamic progression of different classes of population with individuals interacting pairwise. This interaction follows a binary mechanism typical of kinetic theory, wherein agents aim to improve their condition with respect to a mutual utility target. To this end, we introduce kinetic equations of Boltzmann-type to describe the time evolution of the probability distributions of the multi-agent system. The interactions between agents are defined using principles from price theory, specifically employing Cobb-Douglas utility functions for binary exchange and the Edgeworth box to depict the common exchange area where utility increases for both agents. Several numerical experiments presented in the paper highlight the significance of this mechanism in driving the phenomenon toward endemicity. 2023-12-31T12:46:08Z Math. Biosci. Eng. 21 (2024) 4241-4268 Giulia Bertaglia Lorenzo Pareschi Giuseppe Toscani 10.3934/mbe.2024187 http://arxiv.org/abs/2401.00498v3 Analytical Model for Atomic Relaxation in Twisted Moiré Materials 2024-12-31T16:01:31Z By virtue of being atomically thin, the electronic properties of heterostructures built from two-dimensional materials are strongly influenced by atomic relaxation. The atomic layers behave as flexible membranes rather than rigid crystals. Here we develop an analytical theory of lattice relaxation in twisted moiré materials. We obtain analytical results for the lattice displacements and corresponding pseudo gauge fields, as a function of twist angle. We benchmark our results for twisted bilayer graphene and twisted WSe$_2$ bilayers using large-scale molecular dynamics simulations. Our \textit{single-parameter} theory is valid in graphene bilayers for twist angles $θ~\gtrsim 0.7^\circ$, and in twisted WSe$_2$ for $θ~\gtrsim 1.6^\circ$. We also investigate how relaxation alters the electronic structure in twisted bilayer graphene, providing a simple extension to the continuum model to account for lattice relaxation. 2023-12-31T13:37:17Z Phys. Rev. Lett. 133, 266201 (2024) Mohammed M. Al Ezzi Gayani N. Pallewela Christophe De Beule E. J. Mele Shaffique Adam 10.1103/PhysRevLett.133.266201 http://arxiv.org/abs/2401.00500v4 Deformation Quantization with Separation of Variables of $G_{2,4}(\mathbb{C})$ 2025-07-23T05:16:29Z We construct a deformation quantization with separation of variables of the Grassmannian $G_{2,4}(\mathbb{C})$. A star product on $G_{2,4}(\mathbb{C})$ can be explicitly determined as the solution of the recurrence relations for $G_{2,4}(\mathbb{C})$ given by Hara and one of the authors (A. Sako). To provide the solution to the recurrence relations, it is necessary to solve a system of linear equations in each order. However, to give a concrete expression of the general term is not simple because the variables increase with the order of the differentiation of the star product. For this reason, there has been no formula to express the general term of the recurrence relations. In this paper, we overcome this problem by transforming the recurrence relations into simpler ones. We solve the recurrence relations using creation and annihilation operators on a Fock space. From this solution, we obtain an explicit formula of a star product with separation of variables on $G_{2,4}(\mathbb{C})$. 2023-12-31T13:46:54Z SIGMA 21 (2025), 061, 32 pages Taika Okuda Akifumi Sako 10.3842/SIGMA.2025.061 http://arxiv.org/abs/2401.00454v2 Quantum and Classical Communication Complexity of Permutation-Invariant Functions 2025-10-13T15:29:16Z This paper gives a nearly tight characterization of the quantum communication complexity of the permutation-invariant Boolean functions. With such a characterization, we show that the quantum and randomized communication complexity of the permutation-invariant Boolean functions are quadratically equivalent (up to a logarithmic factor). Our results extend a recent line of research regarding query complexity \cite{AA14, Cha19, BCG+20} to communication complexity, showing symmetry prevents exponential quantum speedups. Furthermore, we show the Log-rank Conjecture holds for any non-trivial total permutation-invariant Boolean function. Moreover, we establish a relationship between the quantum/classical communication complexity and the approximate rank of permutation-invariant Boolean functions. This implies the correctness of the Log-approximate-rank Conjecture for permutation-invariant Boolean functions in both randomized and quantum settings (up to a logarithmic factor). 2023-12-31T11:07:49Z reference correction Ziyi Guan Yunqi Huang Penghui Yao Zekun Ye http://arxiv.org/abs/2401.00436v6 Diff-PCR: Diffusion-Based Correspondence Searching in Doubly Stochastic Matrix Space for Point Cloud Registration 2026-04-09T06:05:55Z Efficiently identifying accurate correspondences between point clouds is crucial for both rigid and non-rigid point cloud registration. Existing methods usually rely on geometric or semantic feature embeddings to establish correspondences and then estimate transformations or flow fields. Recently, several state-of-the-art methods have adopted RAFT-like iterative updates to refine solutions. However, these methods still have two major limitations. First, their iterative refinement mechanism lacks transparency, and the update trajectory is largely fixed once the refinement starts, which may lead to suboptimal solutions. Second, they overlook the importance of explicitly refining the correspondence matrix before solving for transformations or flow fields. Most existing approaches compute candidate correspondences in feature space and project the resulting matching matrix only once by using Sinkhorn or dual-softmax normalization. Such a one-shot projection can be far from the globally optimal solution, and these methods usually do not model the distribution of the target matching matrix. In this paper, we propose a novel framework that exploits a denoising diffusion model to predict a search gradient for the optimal matching matrix in doubly stochastic matrix space. Specifically, the diffusion model learns a denoising direction, and the reverse denoising process iteratively searches for improved solutions along this learned direction, which approximates the maximum-likelihood direction of the target matching matrix. To improve efficiency, we design a lightweight denoising module and adopt the accelerated sampling strategy of the Denoising Diffusion Implicit Model (DDIM)\cite{song2020denoising}. Experimental results on 3DMatch/3DLoMatch and 4DMatch/4DLoMatch demonstrate the effectiveness of the proposed framework. 2023-12-31T09:24:28Z Haihua Shi Qianliang Wu http://arxiv.org/abs/2401.00406v1 Low-cost Geometry-based Eye Gaze Detection using Facial Landmarks Generated through Deep Learning 2023-12-31T05:45:22Z Introduction: In the realm of human-computer interaction and behavioral research, accurate real-time gaze estimation is critical. Traditional methods often rely on expensive equipment or large datasets, which are impractical in many scenarios. This paper introduces a novel, geometry-based approach to address these challenges, utilizing consumer-grade hardware for broader applicability. Methods: We leverage novel face landmark detection neural networks capable of fast inference on consumer-grade chips to generate accurate and stable 3D landmarks of the face and iris. From these, we derive a small set of geometry-based descriptors, forming an 8-dimensional manifold representing the eye and head movements. These descriptors are then used to formulate linear equations for predicting eye-gaze direction. Results: Our approach demonstrates the ability to predict gaze with an angular error of less than 1.9 degrees, rivaling state-of-the-art systems while operating in real-time and requiring negligible computational resources. Conclusion: The developed method marks a significant step forward in gaze estimation technology, offering a highly accurate, efficient, and accessible alternative to traditional systems. It opens up new possibilities for real-time applications in diverse fields, from gaming to psychological research. 2023-12-31T05:45:22Z Esther Enhui Ye John Enzhou Ye Joseph Ye Jacob Ye Runzhou Ye http://arxiv.org/abs/2401.00409v1 A Two-stream Hybrid CNN-Transformer Network for Skeleton-based Human Interaction Recognition 2023-12-31T06:46:46Z Human Interaction Recognition is the process of identifying interactive actions between multiple participants in a specific situation. The aim is to recognise the action interactions between multiple entities and their meaning. Many single Convolutional Neural Network has issues, such as the inability to capture global instance interaction features or difficulty in training, leading to ambiguity in action semantics. In addition, the computational complexity of the Transformer cannot be ignored, and its ability to capture local information and motion features in the image is poor. In this work, we propose a Two-stream Hybrid CNN-Transformer Network (THCT-Net), which exploits the local specificity of CNN and models global dependencies through the Transformer. CNN and Transformer simultaneously model the entity, time and space relationships between interactive entities respectively. Specifically, Transformer-based stream integrates 3D convolutions with multi-head self-attention to learn inter-token correlations; We propose a new multi-branch CNN framework for CNN-based streams that automatically learns joint spatio-temporal features from skeleton sequences. The convolutional layer independently learns the local features of each joint neighborhood and aggregates the features of all joints. And the raw skeleton coordinates as well as their temporal difference are integrated with a dual-branch paradigm to fuse the motion features of the skeleton. Besides, a residual structure is added to speed up training convergence. Finally, the recognition results of the two branches are fused using parallel splicing. Experimental results on diverse and challenging datasets, demonstrate that the proposed method can better comprehend and infer the meaning and context of various actions, outperforming state-of-the-art methods. 2023-12-31T06:46:46Z Ruoqi Yin Jianqin Yin http://arxiv.org/abs/2401.00414v1 Is It Possible to Backdoor Face Forgery Detection with Natural Triggers? 2023-12-31T07:16:10Z Deep neural networks have significantly improved the performance of face forgery detection models in discriminating Artificial Intelligent Generated Content (AIGC). However, their security is significantly threatened by the injection of triggers during model training (i.e., backdoor attacks). Although existing backdoor defenses and manual data selection can mitigate those using human-eye-sensitive triggers, such as patches or adversarial noises, the more challenging natural backdoor triggers remain insufficiently researched. To further investigate natural triggers, we propose a novel analysis-by-synthesis backdoor attack against face forgery detection models, which embeds natural triggers in the latent space. We thoroughly study such backdoor vulnerability from two perspectives: (1) Model Discrimination (Optimization-Based Trigger): we adopt a substitute detection model and find the trigger by minimizing the cross-entropy loss; (2) Data Distribution (Custom Trigger): we manipulate the uncommon facial attributes in the long-tailed distribution to generate poisoned samples without the supervision from detection models. Furthermore, to completely evaluate the detection models towards the latest AIGC, we utilize both state-of-the-art StyleGAN and Stable Diffusion for trigger generation. Finally, these backdoor triggers introduce specific semantic features to the generated poisoned samples (e.g., skin textures and smile), which are more natural and robust. Extensive experiments show that our method is superior from three levels: (1) Attack Success Rate: ours achieves a high attack success rate (over 99%) and incurs a small model accuracy drop (below 0.2%) with a low poisoning rate (less than 3%); (2) Backdoor Defense: ours shows better robust performance when faced with existing backdoor defense methods; (3) Human Inspection: ours is less human-eye-sensitive from a comprehensive user study. 2023-12-31T07:16:10Z Xiaoxuan Han Songlin Yang Wei Wang Ziwen He Jing Dong http://arxiv.org/abs/2401.00444v1 RIS-Enabled Integrated Sensing and Communication for 6G Systems 2023-12-31T10:10:22Z The following paper proposes a new target localization system design using an architecture based on reconfigurable intelligent surfaces (RISs) and passive radars (PRs) for integrated sensing and communications systems. The preamble of the communication signal is exploited in order to perform target sensing tasks, which involve detection and localization. The RIS in this case can aid the PR in sensing targets that are otherwise not seen by the PR itself, due to the many obstacles encountered within the propagation channel. Therefore, this work proposes a localization algorithm tailored for the integrated sensing and communications RIS-aided architecture, which is capable of uniquely positioning targets within the scene. The algorithm is capable of detecting the number of targets along with estimating the position of targets via angles and times of arrival. Our simulation results demonstrate the performance of the localization method in terms of different localization and detection metrics and for increasing RIS sizes. 2023-12-31T10:10:22Z IEEE Wireless Communications and Networking Conference, 2024 Dexin Wang Ahmad Bazzi Marwa Chafii http://arxiv.org/abs/2401.00447v1 User Clustering for STAR-RIS Assisted Full-Duplex NOMA Communication Systems 2023-12-31T10:33:34Z In contrast to conventional reconfigurable intelligent surface (RIS), simultaneous transmitting and reflecting reconfigurable intelligent surface (STAR-RIS) has been proposed recently to enlarge the serving area from 180o to 360o coverage. This work considers the performance of a STAR-RIS aided full-duplex (FD) non-orthogonal multiple access (NOMA) communication systems. The STAR-RIS is implemented at the cell-edge to assist the cell-edge users, while the cell-center users can communicate directly with a FD base station (BS). We first introduce new user clustering schemes for the downlink and uplink transmissions. Then, based on the proposed transmission schemes closed-form expressions of the ergodic rates in the downlink and uplink modes are derived taking into account the system impairments caused by the self interference at the FD-BS and the imperfect successive interference cancellation (SIC). Moreover, an optimization problem to maximize the total sum-rate is formulated and solved by optimizing the amplitudes and the phase-shifts of the STAR-RIS elements and allocating the transmit power efficiently. The performance of the proposed user clustering schemes and the optimal STAR-RIS design are investigated through numerical results 2023-12-31T10:33:34Z arXiv admin note: text overlap with arXiv:2309.15037 Abdelhamid Salem Kai-Kit Wong Chan-Byoung Chae Yangyang Zhang http://arxiv.org/abs/2401.00463v2 Analyzing Local Representations of Self-supervised Vision Transformers 2024-03-21T14:57:25Z In this paper, we present a comparative analysis of various self-supervised Vision Transformers (ViTs), focusing on their local representative power. Inspired by large language models, we examine the abilities of ViTs to perform various computer vision tasks with little to no fine-tuning. We design evaluation framework to analyze the quality of local, i.e.\ patch-level, representations in the context of few-shot semantic segmentation, instance identification, object retrieval and tracking. We discover that contrastive learning based methods like DINO produce more universal patch representations that can be immediately applied for downstream tasks with no parameter tuning, compared to masked image modeling. The embeddings learned using the latter approach, e.g. in masked autoencoders, have high variance features that harm distance-based algorithms, such as k-NN, and do not contain useful information for most downstream tasks. Furthermore, we demonstrate that removing these high-variance features enhances k-NN for MAE, as well as for its recent extension Scale-MAE. Finally, we find an object instance retrieval setting where DINOv2, a model pretrained on two orders of magnitude more data, falls short of its less compute intensive counterpart DINO. 2023-12-31T11:38:50Z Ani Vanyan Alvard Barseghyan Hakob Tamazyan Vahan Huroyan Hrant Khachatrian Martin Danelljan http://arxiv.org/abs/2401.00467v1 Amplification of femtosecond pulses with AI-assisted spectral phase modulation 2023-12-31T11:45:14Z We report our investigation on ultrashort laser pulse optimization using an AI algorithm in a system consisting of a mode-locked oscillator, a spectral phase shaper, and a highly nonlinear amplifier. We analyzed the performance of the pulse optimization process as a function of two main parameters: the resolution of spectral phase modulation and the number of agents in the algorithm. We showed that the algorithm could find an optimum phase profile for the seed pulse, which allowed for a reduction of the FWHM of the amplified pulse by 10 fs (from 46 to 36 fs), and significantly reduced the intensity of the side-pulse by a factor of 4.6. Importantly, the algorithm used does not require any training and optimizes the pulse shape without any knowledge about the input pulse parameters or the parameters of the amplifier. We believe the proposed system might be a convenient test bed for evaluating various AI-based algorithms in a pulse optimization task. 2023-12-31T11:45:14Z 10 pages, 8 figures Mikołaj Krakowski Alicja Kwaśny Grzegorz Soboń http://arxiv.org/abs/2401.00468v1 Blockchain and Deep Learning-Based IDS for Securing SDN-Enabled Industrial IoT Environments 2023-12-31T11:49:42Z The industrial Internet of Things (IIoT) involves the integration of Internet of Things (IoT) technologies into industrial settings. However, given the high sensitivity of the industry to the security of industrial control system networks and IIoT, the use of software-defined networking (SDN) technology can provide improved security and automation of communication processes. Despite this, the architecture of SDN can give rise to various security threats. Therefore, it is of paramount importance to consider the impact of these threats on SDN-based IIoT environments. Unlike previous research, which focused on security in IIoT and SDN architectures separately, we propose an integrated method including two components that work together seamlessly for better detecting and preventing security threats associated with SDN-based IIoT architectures. The two components consist in a convolutional neural network-based Intrusion Detection System (IDS) implemented as an SDN application and a Blockchain-based system (BS) to empower application layer and network layer security, respectively. A significant advantage of the proposed method lies in jointly minimizing the impact of attacks such as command injection and rule injection on SDN-based IIoT architecture layers. The proposed IDS exhibits superior classification accuracy in both binary and multiclass categories. 2023-12-31T11:49:42Z Samira Kamali Poorazad Chafika Benzaıd Tarik Taleb http://arxiv.org/abs/2401.00476v1 Reproductive outcome in female wistar rats treated with nhexane, dichloromethane and aqueous ethanol extracts of Cucurbita pepo seed 2023-12-31T12:32:05Z In developing countries, healthcare challenges and expensive infertility treatments has resulted in resurgent interest in medicinal plants. This study was designed to determine if Curcubita pepo seed can enhance female fertility, by assessing the reproductive outcome in female wistar rats treated with n-hexane (nHE), dichloromethane (DCM) and aqueous ethanol (Aq. Eth) extracts of Curcubita pepo seed. Total of 48 rats randomly grouped into 12 (n=4), were treated for 21 days by oral gavage as follows: A (control) = 0.5ml 20% tween 80 (vehicle); B (positive control) = 10mg/kg clomiphene citrate, C, D & E = 142.86, 285.71 and 428.57 mg/kg nHE; F, G & H = 142.86, 285.71 and 428.57 mg/kg DCM ; and I, J & K =142.86, 285.71 and 428.57 mg/kg Aq.Eth extracts. Group L (positive control 2) = 10mg/kg clomiphene citrate for 8 days. Following treatment, the rats were paired with males for mating, designating the confirmation day as gestational day 0 (GD 0). On GD 20, the animals were laparatomised and reproductive outcome was determined by assessing foetal weight, foetal crown-rump length, litter size, number of implantation and resorption sites. Results showed all extracts had no significant (p >0.05) effect on the reproductive outcome indices. Clomiphene citrate significantly decreased reproductive outcome indices. In conclusion, Cucurbita pepo seed did not enhance the reproductive outcome of treated female rats at the doses and duration used in this study. This finding may serve as a springboard for future studies exploring the effect of C.pepo at different doses or durations. 2023-12-31T12:32:05Z 5 tables, 2 figures Anyanwu C. F Georgewill O. A. Obinna Victoria C http://arxiv.org/abs/2401.00490v2 Kernel Density Estimation for Multiclass Quantification 2024-01-02T19:52:24Z Several disciplines, like the social sciences, epidemiology, sentiment analysis, or market research, are interested in knowing the distribution of the classes in a population rather than the individual labels of the members thereof. Quantification is the supervised machine learning task concerned with obtaining accurate predictors of class prevalence, and to do so particularly in the presence of label shift. The distribution-matching (DM) approaches represent one of the most important families among the quantification methods that have been proposed in the literature so far. Current DM approaches model the involved populations by means of histograms of posterior probabilities. In this paper, we argue that their application to the multiclass setting is suboptimal since the histograms become class-specific, thus missing the opportunity to model inter-class information that may exist in the data. We propose a new representation mechanism based on multivariate densities that we model via kernel density estimation (KDE). The experiments we have carried out show our method, dubbed KDEy, yields superior quantification performance with respect to previous DM approaches. We also investigate the KDE-based representation within the maximum likelihood framework and show KDEy often shows superior performance with respect to the expectation-maximization method for quantification, arguably the strongest contender in the quantification arena to date. 2023-12-31T13:19:27Z fixed broken references to appendices Alejandro Moreo Pablo González Juan José del Coz http://arxiv.org/abs/2401.00494v1 Generalization of the Bargmann-Wigner approach to constructing relativistic fields 2023-12-31T13:28:28Z We review the method for constructing local relativistic fields corresponding to the Bargmann-Wigner wave functions that describe the unitary irreducible representations of the $4D$ Poincaré group. The method is based on the use of the generalized Wigner operator connecting the wave functions of induced representations and local relativistic fields. Applications of this operator for constructing massive local relativistic fields as well as massless helicity local fields and massless local infinite spin fields are considered. 2023-12-31T13:28:28Z 1+12 pages, Contribution to the Proceedings of the International Conference on Particle Physics and Cosmology (professor V.A. Rubakov memorial conference), October 02-07, 2023, Yerevan, Armenia I. L. Buchbinder S. A. Fedoruk A. P. Isaev M. A. Podoinitsyn http://arxiv.org/abs/2401.00495v1 Lattice construction of mixed 't Hooft anomaly with higher-form symmetry 2023-12-31T13:30:48Z In this talk, we give the lattice regularized formulation of the mixed 't Hooft anomaly between the $\mathbb{Z}_N$ $1$-form symmetry and the $θ$ periodicity for $4$d pure Yang-Mills theory, which was originally discussed by Gaiotto $\textit{et al.}$ in the continuum description. For this purpose, we define the topological charge of the lattice $SU(N)$ gauge theory coupled with the background $\mathbb{Z}_N$ $2$-form gauge fields $B_p$ by generalizing Lüscher's construction of the $SU(N)$ topological charge. We show that this lattice topological charge enjoys the fractional $1/N$ shift completely characterized by the background gauge field $B_p$, and this rigorously proves the mixed 't Hooft anomaly with the finite lattice spacings. As a consequence, the Yang-Mills vacua at $θ$ and $θ+2π$ are distinct as the symmetry-protected topological states when the confinement is assumed. 2023-12-31T13:30:48Z 8 pages, 2 figures, talk presented at the 40th International Symposium on Lattice Field Theory (Lattice2023), July 31st - August 4th, 2023, Fermi National Accelerator Laboratory Motokazu Abe Okuto Morikawa Soma Onoda Hiroshi Suzuki Yuya Tanizaki http://arxiv.org/abs/2401.00428v3 Training toward significance with the decorrelated event classifier transformer neural network 2024-07-11T01:50:17Z Experimental particle physics uses machine learning for many tasks, where one application is to classify signal and background events. This classification can be used to bin an analysis region to enhance the expected significance for a mass resonance search. In natural language processing, one of the leading neural network architectures is the transformer. In this work, an event classifier transformer is proposed to bin an analysis region, in which the network is trained with special techniques. The techniques developed here can enhance the significance and reduce the correlation between the network's output and the reconstructed mass. It is found that this trained network can perform better than boosted decision trees and feed-forward networks. 2023-12-31T08:57:29Z 11 pages, 7 figures, 1 table Phys. Rev. D 109, 096035 (2024) Jaebak Kim http://arxiv.org/abs/2401.00412v2 Toward the theoretically observable limit of electron density distribution by single-crystal synchrotron X-ray diffraction: The case of orbitally ordered Ti-3d^1 in YTiO_3 2024-07-08T16:43:10Z The theoretically observable limit of electron density distribution by single-crystal X-ray diffraction is discussed. When F_{orb} and δF are defined as, respectively, the partial structure factor for an orbital and the deviation of the observed F from the true F, the accuracy of electron density attributable to F_{orb} is chiefly determined by the number of reflections satisfying the condition F_{orb}/F > δF/F. Since F_{orb}/F, which is generally small for crystals with large F(0,0,0), is constant under a given set of experimental conditions, δF/F must be reduced to increase the number of reflections satisfying F_{orb}/F > δF/F. The present paper demonstrates how to reduce δF mathematically and experimentally, and the following topics are covered: the Poisson statistics, accumulation of errors in the data collection and reduction procedure, multiple diffraction, conversion error from F^2 to F in refinement programs, which is unavoidable when the input quantities have different dimension from F, weighting of reflections, and tips. For demonstration, observation of the electron density of the Ti-3d^1 orbital in YTiO_3 by synchrotron single-crystal X-ray diffraction is presented. 2023-12-31T07:08:06Z 68 pages, 20 figures Terutoshi Sakakura Yoshihisa Ishikawa Shunji Kishimoto Yasuyuki Takenaka Kiyoaki Tanaka Shigeki Miyasaka Yoshinori Tokura Yukio Noda Nobuo Ishizawa Hajime Sagayama Hajime Yamamoto Hiroyuki Kimura http://arxiv.org/abs/2401.00477v2 Coding for Gaussian Two-Way Channels: Linear and Learning-Based Approaches 2025-04-23T13:16:13Z Although user cooperation cannot improve the capacity of Gaussian two-way channels (GTWCs) with independent noises, it can improve communication reliability. In this work, we aim to enhance and balance the communication reliability in GTWCs by minimizing the sum of error probabilities via joint design of encoders and decoders at the users. We first formulate general encoding/decoding functions, where the user cooperation is captured by the coupling of user encoding processes. The coupling effect renders the encoder/decoder design non-trivial, requiring effective decoding to capture this effect, as well as efficient power management at the encoders within power constraints. To address these challenges, we propose two different two-way coding strategies: linear coding and learning-based coding. For linear coding, we propose optimal linear decoding and discuss new insights on encoding regarding user cooperation to balance reliability. We then propose an efficient algorithm for joint encoder/decoder design. For learning-based coding, we introduce a novel recurrent neural network (RNN)-based coding architecture, where we propose interactive RNNs and a power control layer for encoding, and we incorporate bi-directional RNNs with an attention mechanism for decoding. Through simulations, we show that our two-way coding methodologies outperform conventional channel coding schemes (that do not utilize user cooperation) significantly in sum-error performance. We also demonstrate that our linear coding excels at high signal-to-noise ratios (SNRs), while our RNN-based coding performs best at low SNRs. We further investigate our two-way coding strategies in terms of power distribution, two-way coding benefit, different coding rates, and block-length gain. 2023-12-31T12:40:18Z This work has been accepted for publication in the IEEE Transactions on Information Theory Junghoon Kim Taejoon Kim Anindya Bijoy Das Seyyedali Hosseinalipour David J. Love Christopher G. Brinton http://arxiv.org/abs/2401.00473v1 Emulating insect brains for neuromorphic navigation 2023-12-31T12:05:42Z Bees display the remarkable ability to return home in a straight line after meandering excursions to their environment. Neurobiological imaging studies have revealed that this capability emerges from a path integration mechanism implemented within the insect's brain. In the present work, we emulate this neural network on the neuromorphic mixed-signal processor BrainScaleS-2 to guide bees, virtually embodied on a digital co-processor, back to their home location after randomly exploring their environment. To realize the underlying neural integrators, we introduce single-neuron spike-based short-term memory cells with axo-axonic synapses. All entities, including environment, sensory organs, brain, actuators, and the virtual body, run autonomously on a single BrainScaleS-2 microchip. The functioning network is fine-tuned for better precision and reliability through an evolution strategy. As BrainScaleS-2 emulates neural processes 1000 times faster than biology, 4800 consecutive bee journeys distributed over 320 generations occur within only half an hour on a single neuromorphic core. 2023-12-31T12:05:42Z Korbinian Schreiber Timo Wunderlich Philipp Spilger Sebastian Billaudelle Benjamin Cramer Yannik Stradmann Christian Pehle Eric Müller Mihai A. Petrovici Johannes Schemmel Karlheinz Meier http://arxiv.org/abs/2401.00457v4 Two types of filtrations for $\mathrm{wK4}$ and its relatives 2025-06-12T18:14:15Z We study the finite model property of subframe logics with expressible transitive reflexive closure modality. For $m>0$, let $\mathrm{L}_m$ be the logic defined by axiom $\lozenge^{m+1} p\to \lozenge p\vee p$. We construct filtrations for the logics $\mathrm{L}_m$. It follows that these logics and their tense counterparts have the finite model property. Then we show that every canonical subframe logic that contains $\mathrm{L}_m$ have the finite model property. 2023-12-31T11:18:59Z Andrey Kudinov Ilya Shapirovsky http://arxiv.org/abs/2401.00413v2 Real-Time FJ/MAC PDE Solvers via Tensorized, Back-Propagation-Free Optical PINN Training 2024-01-04T06:25:16Z Solving partial differential equations (PDEs) numerically often requires huge computing time, energy cost, and hardware resources in practical applications. This has limited their applications in many scenarios (e.g., autonomous systems, supersonic flows) that have a limited energy budget and require near real-time response. Leveraging optical computing, this paper develops an on-chip training framework for physics-informed neural networks (PINNs), aiming to solve high-dimensional PDEs with fJ/MAC photonic power consumption and ultra-low latency. Despite the ultra-high speed of optical neural networks, training a PINN on an optical chip is hard due to (1) the large size of photonic devices, and (2) the lack of scalable optical memory devices to store the intermediate results of back-propagation (BP). To enable realistic optical PINN training, this paper presents a scalable method to avoid the BP process. We also employ a tensor-compressed approach to improve the convergence and scalability of our optical PINN training. This training framework is designed with tensorized optical neural networks (TONN) for scalable inference acceleration and MZI phase-domain tuning for \textit{in-situ} optimization. Our simulation results of a 20-dim HJB PDE show that our photonic accelerator can reduce the number of MZIs by a factor of $1.17\times 10^3$, with only $1.36$ J and $1.15$ s to solve this equation. This is the first real-size optical PINN training framework that can be applied to solve high-dimensional PDEs. 2023-12-31T07:10:15Z ML with New Compute Paradigms (MLNCP) at NeurIPS 2023 Yequan Zhao Xian Xiao Xinling Yu Ziyue Liu Zhixiong Chen Geza Kurczveil Raymond G. Beausoleil Zheng Zhang http://arxiv.org/abs/2401.00417v2 Stability for the 2-D plane Poiseuille flow in finite channel 2024-03-03T04:51:57Z In this paper, we study the stability for 2-D plane Poiseuille flow $(1-y^2,0)$ in a channel $\mathbb{T}\times (-1,1)$ with Navier-slip boundary condition. We prove that if the initial perturbation for velocity field $u_0$ satisfies that $\|u_0\|_{H^{\frac{7}{2}+}} \leq ε_1 ν^{2/3}$ for some suitable small $0<ε_1 \ll 1$ independent of viscosity coefficient $ν$, then the solution to the Navier-Stokes equations is global in time and does not transit from the plane Poiseuille flow. This result improves the result of \cite{DL1} from $3/4$ to $2/3$. 2023-12-31T07:44:58Z This version fixes an error in the proof of precious version, and improves the result of [18] form 3/4 to 2/3 for slip boundary value problem. The case of non-slip boundary value problem is not included in this version Shijin Ding Zhilin Lin http://arxiv.org/abs/2401.00423v1 MSGNet: Learning Multi-Scale Inter-Series Correlations for Multivariate Time Series Forecasting 2023-12-31T08:23:24Z Multivariate time series forecasting poses an ongoing challenge across various disciplines. Time series data often exhibit diverse intra-series and inter-series correlations, contributing to intricate and interwoven dependencies that have been the focus of numerous studies. Nevertheless, a significant research gap remains in comprehending the varying inter-series correlations across different time scales among multiple time series, an area that has received limited attention in the literature. To bridge this gap, this paper introduces MSGNet, an advanced deep learning model designed to capture the varying inter-series correlations across multiple time scales using frequency domain analysis and adaptive graph convolution. By leveraging frequency domain analysis, MSGNet effectively extracts salient periodic patterns and decomposes the time series into distinct time scales. The model incorporates a self-attention mechanism to capture intra-series dependencies, while introducing an adaptive mixhop graph convolution layer to autonomously learn diverse inter-series correlations within each time scale. Extensive experiments are conducted on several real-world datasets to showcase the effectiveness of MSGNet. Furthermore, MSGNet possesses the ability to automatically learn explainable multi-scale inter-series correlations, exhibiting strong generalization capabilities even when applied to out-of-distribution samples. 2023-12-31T08:23:24Z 13 pages, 12 figures Wanlin Cai Yuxuan Liang Xianggen Liu Jianshuai Feng Yuankai Wu http://arxiv.org/abs/2401.00427v2 The functional volume product under heat flow 2024-03-20T14:44:53Z We prove that the functional volume product for even functions is monotone increasing along the Fokker--Planck heat flow. This in particular yields a new proof of the functional Blaschke--Santaló inequality by K. Ball and also Artstein-Avidan--Klartag--Milman in the even case. This result is the consequence of a new understanding of the regularizing property of the Ornstein--Uhlenbeck semigroup. That is, we establish an improvement of Borell's reverse hypercontractivity inequality for even functions and identify the sharp range of the admissible exponents. As another consequence of successfully identifying the sharp range for the inequality, we derive the sharp $L^p$-$L^q$ inequality for the Laplace transform for even functions. The best constant of the inequality is attained by centered Gaussians, and thus this provides an analogous result to Beckner's sharp Hausdorff--Young inequality. Our technical novelty in the proof is the use of the Brascamp--Lieb inequality for log-concave measures and Cramér--Rao's inequality in this context. 2023-12-31T08:48:33Z In this update, we have mentioned about the "detropicalised" approach that has been proposed in the discussion of Klartag and Tao in Tao's blog post as it is closely related this work. We have mentioned works of Berndtsson--Mastrantonis--Rubinstein and Kolesnikov--Werner. Also, we have split the result on the stability from this version. This part will be in the forthcoming paper Shohei Nakamura Hiroshi Tsuji http://arxiv.org/abs/2401.00452v3 Multi-scale cross-attention transformer encoder for event classification 2024-02-15T02:43:00Z We deploy an advanced Machine Learning (ML) environment, leveraging a multi-scale cross-attention encoder for event classification, towards the identification of the $gg\to H\to hh\to b\bar b b\bar b$ process at the High Luminosity Large Hadron Collider (HL-LHC), where $h$ is the discovered Standard Model (SM)-like Higgs boson and $H$ a heavier version of it (with $m_H>2m_h$). In the ensuing boosted Higgs regime, the final state consists of two fat jets. Our multi-modal network can extract information from the jet substructure and the kinematics of the final state particles through self-attention transformer layers. The diverse learned information is subsequently integrated to improve classification performance using an additional transformer encoder with cross-attention heads. We ultimately prove that our approach surpasses in performance current alternative methods used to establish sensitivity to this process, whether solely based on kinematic analysis or else on a combination of this with mainstream ML approaches. Then, we employ various interpretive methods to evaluate the network results, including attention map analysis and visual representation of Gradient-weighted Class Activation Mapping (Grad-CAM). Finally, we note that the proposed network is generic and can be applied to analyse any process carrying information at different scales. Our code is publicly available for generic use. 2023-12-31T11:03:28Z Typos corrected A. Hammad S. Moretti M. Nojiri http://arxiv.org/abs/2401.00458v1 Two sequences of spiral galaxies with different shapes of the metallicity gradients 2023-12-31T11:20:27Z We considered two sequences of spiral galaxies with different shapes of the radial gas-phase oxygen abundance distributions from the galaxies in the MaNGA survey: (1) Galaxies in which the gradient is well approximated by a single linear relation across the whole disc, that is, galaxies with an S (slope) gradients, (2) galaxies in which the metallicity in the inner region of the disc is at a nearly constant level and the gradient is negative at larger radii, that is, galaxies with level-slope (LS) gradients. We also selected galaxies with a nearly uniform oxygen abundance across the whole galaxy, that is, galaxies with level (L) gradients that can be the final evolutionary stage of the two galaxy sequences described above. The radial nitrogen abundance distributions in galaxies with LS oxygen abundance distributions also show breaks at radii smaller than the O/H distribution breaks. The observed behaviour of the oxygen and nitrogen abundances with radius in these galaxies can be explained by the time delay between the nitrogen and oxygen enrichment together with the variation in the star formation history along the radius. These galaxies clearly show the effect of the inside-out disc evolution model. We find that the shape of the radial abundance distribution in a galaxy is not related to its macroscopic characteristics (rotation velocity, stellar mass, isophotal radius, and star formation rate). The correlations between the gradient slopes and macroscopic characteristics of galaxies are weak in the sense that the scatter of the points in each diagram is large. We also examined the properties of the Milky Way in the context of the considered galaxy samples. 2023-12-31T11:20:27Z 20 pages, 19 figures, accepted to Astronomy and Astrophysics L. S. Pilyugin G. Tautvaisiene http://arxiv.org/abs/2401.00485v1 Additive spectrum preserving mappings from von Neumann algebras 2023-12-31T12:57:45Z We establish Jafarian's 2009 conjecture that every additive spectrum preserving mapping from a von Neumann algebra onto a semisimple Banach algebra is a Jordan isomorphism. 2023-12-31T12:57:45Z 13 pages Martin Mathieu Francois Schulz http://arxiv.org/abs/2401.00499v3 Generating High-Precision Force Fields for Molecular Dynamics Simulations to Study Chemical Reaction Mechanisms using Molecular Configuration Transformer 2024-04-11T17:15:43Z Theoretical studies on chemical reaction mechanisms have been crucial in organic chemistry. Traditionally, calculating the manually constructed molecular conformations of transition states for chemical reactions using quantum chemical calculations is the most commonly used method. However, this way is heavily dependent on individual experience and chemical intuition. In our previous study, we proposed a research paradigm that uses enhanced sampling in molecular dynamics simulations to study chemical reactions. This approach can directly simulate the entire process of a chemical reaction. However, the computational speed limits the use of high-precision potential energy functions for simulations. To address this issue, we present a scheme for training high-precision force fields for molecular modeling using a previously developed graph-neural-network-based molecular model, molecular configuration transformer. This potential energy function allows for highly accurate simulations at a low computational cost, leading to more precise calculations of the mechanism of chemical reactions. We applied this approach to study a Claisen rearrangement reaction and a Carbonyl insertion reaction catalyzed by Manganese. 2023-12-31T13:43:41Z Sihao Yuan Xu Han Jun Zhang Zhaoxin Xie Cheng Fan Yunlong Xiao Yi Qin Gao Yi Isaac Yang http://arxiv.org/abs/2401.00435v1 Bidirectional Trained Tree-Structured Decoder for Handwritten Mathematical Expression Recognition 2023-12-31T09:24:21Z The Handwritten Mathematical Expression Recognition (HMER) task is a critical branch in the field of OCR. Recent studies have demonstrated that incorporating bidirectional context information significantly improves the performance of HMER models. However, existing methods fail to effectively utilize bidirectional context information during the inference stage. Furthermore, current bidirectional training methods are primarily designed for string decoders and cannot adequately generalize to tree decoders, which offer superior generalization capabilities and structural analysis capacity. In order to overcome these limitations, we propose the Mirror-Flipped Symbol Layout Tree (MF-SLT) and Bidirectional Asynchronous Training (BAT) structure. Our method extends the bidirectional training strategy to the tree decoder, allowing for more effective training by leveraging bidirectional information. Additionally, we analyze the impact of the visual and linguistic perception of the HMER model separately and introduce the Shared Language Modeling (SLM) mechanism. Through the SLM, we enhance the model's robustness and generalization when dealing with visual ambiguity, particularly in scenarios with abundant training data. Our approach has been validated through extensive experiments, demonstrating its ability to achieve new state-of-the-art results on the CROHME 2014, 2016, and 2019 datasets, as well as the HME100K dataset. The code used in our experiments will be publicly available. 2023-12-31T09:24:21Z Hanbo Cheng Chenyu Liu Pengfei Hu Zhenrong Zhang Jiefeng Ma Jun Du http://arxiv.org/abs/2401.00474v1 Probabilistically Checkable Reconfiguration Proofs and Inapproximability of Reconfiguration Problems 2023-12-31T12:22:11Z Motivated by the inapproximability of reconfiguration problems, we present a new PCP-type characterization of PSPACE, which we call a probabilistically checkable reconfiguration proof (PCRP): Any PSPACE computation can be encoded into an exponentially long sequence of polynomially long proofs such that every adjacent pair of the proofs differs in at most one bit, and every proof can be probabilistically checked by reading a constant number of bits. Using the new characterization, we prove PSPACE-completeness of approximate versions of many reconfiguration problems, such as the Maxmin $3$-SAT Reconfiguration problem. This resolves the open problem posed by Ito, Demaine, Harvey, Papadimitriou, Sideri, Uehara, and Uno (ISAAC 2008; Theor. Comput. Sci. 2011) as well as the Reconfiguration Inapproximability Hypothesis by Ohsaka (STACS 2023) affirmatively. We also present PSPACE-completeness of approximating the Maxmin Clique Reconfiguration problem to within a factor of $n^ε$ for some constant $ε> 0$. 2023-12-31T12:22:11Z 31 pages Proceedings of the 56th Annual ACM Symposium on Theory of Computing (STOC), pp. 1435--1445, 2024 Shuichi Hirahara Naoto Ohsaka 10.1145/3618260.3649667 http://arxiv.org/abs/2401.00462v2 On the existence of analytic families of G-stable lattices and their reductions 2024-11-18T23:15:33Z In this article, we prove the existence of rigid analytic families of $G$-stable lattices with locally constant reductions inside families of representations of a topologically compact group $G$, extending a result of Hellman obtained in the semi-simple residual case. Implementing this generalization in the context of Galois representations, we prove a local constancy result for reductions modulo prime powers of trianguline representations of generic dimension $d$. Moreover, we present two explicit applications. First, in dimension two, we extend to a prime power setting and to the whole rigid projective line a recent result of Bergdall, Levin and Liu concerning reductions of semi-stable representations of $\text{Gal}(\overline{\mathbb{Q}}_p / \mathbb{Q}_p)$ with fixed Hodge-Tate weights and large $\mathcal{L}$-invariant. Second, in dimension $d$, let $V_n$ be a sequence of crystalline representations converging in a certain geometric sense to a crystalline representation $V$. We show that for any refined version $(V, σ)$ of $V$ (or equivalently for any chosen triangulation of its attached $(\varphi, Γ)$-module $D_{\text{rig}} (V)$ over the Robba ring), there exists a sequence of refinement $σ_n$ of each of the $V_n$ such that the limit as refined representations $(V_n , σ_n )$ converges to the $(V, σ)$. This result does not hold under the weaker assumption that $V_n$ converges only uniformly $p$-adically to $V$ (in the sense of Chenevier, Khare and Larsen). 2023-12-31T11:37:24Z Emiliano Torti http://arxiv.org/abs/2401.00404v2 Explicit Generators for the Stabilizers of Rational Points in Thompson's Group $F$ 2024-11-20T05:39:38Z We construct explicit finite generating sets for the stabilizers in Thompson's group $F$ of rational points of a unit interval or a Cantor set. Our technique is based on the Reidemeister-Schreier procedure in the context of Schreier graphs of such stabilizers in $F$. It is well known that the stabilizers of dyadic rational points are isomorphic to $F\times F$ and can thus be generated by 4 explicit elements. We show that the stabilizer of every non-dyadic rational point $b\in (0,1)$ is generated by 5 elements that are explicitly calculated as words in generators $x_0, x_1$ of $F$ that depend on the binary expansion of $b$. We also provide an alternative simple proof that the stabilizers of all rational points are finitely presented. 2023-12-31T05:38:29Z 19 pages, 9 figures and pictures Krystofer Baker Dmytro Savchuk http://arxiv.org/abs/2401.00408v2 Computing greatest common divisor of several parametric univariate polynomials via generalized subresultant polynomials 2024-09-06T06:46:15Z In this paper, we tackle the following problem: compute the gcd for several univariate polynomials with parametric coefficients. It amounts to partitioning the parameter space into ``cells'' so that the gcd has a uniform expression over each cell and constructing a uniform expression of gcd in each cell. We tackle the problem as follows. We begin by making a natural and obvious extension of subresultant polynomials of two polynomials to several polynomials. Then we develop the following structural theories about them. 1. We generalize Sylvester's theory to several polynomials, in order to obtain an elegant relationship between generalized subresultant polynomials and the gcd of several polynomials, yielding an elegant algorithm. 2. We generalize Habicht's theory to several polynomials, in order to obtain a systematic relationship between generalized subresultant polynomials and pseudo-remainders, yielding an efficient algorithm. Using the generalized theories, we present a simple (structurally elegant) algorithm which is significantly more efficient (both in the output size and computing time) than algorithms based on previous approaches. 2023-12-31T06:32:54Z Hoon Hong Jing Yang http://arxiv.org/abs/2401.00403v2 Overcome Modal Bias in Multi-modal Federated Learning via Balanced Modality Selection 2024-07-28T14:33:47Z Selecting proper clients to participate in each federated learning (FL) round is critical to effectively harness a broad range of distributed data. Existing client selection methods simply consider the mining of distributed uni-modal data, yet, their effectiveness may diminish in multi-modal FL (MFL) as the modality imbalance problem not only impedes the collaborative local training but also leads to a severe global modality-level bias. We empirically reveal that local training with a certain single modality may contribute more to the global model than training with all local modalities. To effectively exploit the distributed multiple modalities, we propose a novel Balanced Modality Selection framework for MFL (BMSFed) to overcome the modal bias. On the one hand, we introduce a modal enhancement loss during local training to alleviate local imbalance based on the aggregated global prototypes. On the other hand, we propose the modality selection aiming to select subsets of local modalities with great diversity and achieving global modal balance simultaneously. Our extensive experiments on audio-visual, colored-gray, and front-back datasets showcase the superiority of BMSFed over baselines and its effectiveness in multi-modal data exploitation. 2023-12-31T05:37:27Z Accepted by ECCV24, 23 pages Yunfeng Fan Wenchao Xu Haozhao Wang Fushuo Huo Jinyu Chen Song Guo http://arxiv.org/abs/2401.00416v2 SVFAP: Self-supervised Video Facial Affect Perceiver 2024-10-01T07:55:22Z Video-based facial affect analysis has recently attracted increasing attention owing to its critical role in human-computer interaction. Previous studies mainly focus on developing various deep learning architectures and training them in a fully supervised manner. Although significant progress has been achieved by these supervised methods, the longstanding lack of large-scale high-quality labeled data severely hinders their further improvements. Motivated by the recent success of self-supervised learning in computer vision, this paper introduces a self-supervised approach, termed Self-supervised Video Facial Affect Perceiver (SVFAP), to address the dilemma faced by supervised methods. Specifically, SVFAP leverages masked facial video autoencoding to perform self-supervised pre-training on massive unlabeled facial videos. Considering that large spatiotemporal redundancy exists in facial videos, we propose a novel temporal pyramid and spatial bottleneck Transformer as the encoder of SVFAP, which not only largely reduces computational costs but also achieves excellent performance. To verify the effectiveness of our method, we conduct experiments on nine datasets spanning three downstream tasks, including dynamic facial expression recognition, dimensional emotion recognition, and personality recognition. Comprehensive results demonstrate that SVFAP can learn powerful affect-related representations via large-scale self-supervised pre-training and it significantly outperforms previous state-of-the-art methods on all datasets. Code is available at https://github.com/sunlicai/SVFAP. 2023-12-31T07:44:05Z Published in: IEEE Transactions on Affective Computing (Early Access). The code and models are available at https://github.com/sunlicai/SVFAP IEEE Transactions on Affective Computing, 2024 Licai Sun Zheng Lian Kexin Wang Yu He Mingyu Xu Haiyang Sun Bin Liu Jianhua Tao 10.1109/TAFFC.2024.3436913 http://arxiv.org/abs/2401.00450v1 Fault-tolerant quantum computation by hybrid qubits with bosonic cat-code and single photons 2023-12-31T10:57:31Z Hybridizing different degrees of freedom or physical platforms potentially offers various advantages in building scalable quantum architectures. We here introduce a fault-tolerant hybrid quantum computation by taking the advantages of both discrete variable (DV) and continuous variable (CV) systems. Particularly, we define a CV-DV hybrid qubit with bosonic cat-code and single photon, which is implementable in current photonic platforms. By the cat-code encoded in the CV part, the dominant loss errors are readily correctable without multi-qubit encoding, while the logical basis is inherently orthogonal due to the DV part. We design fault-tolerant architectures by concatenating hybrid qubits and an outer DV quantum error correction code such as topological codes, exploring their potential merits in developing scalable quantum computation. We demonstrate by numerical simulations that our scheme is at least an order of magnitude more resource-efficient over all previous proposals in photonic platforms, allowing to achieve a record-high loss threshold among existing CV and hybrid approaches. We discuss its realization not only in all-photonic platforms but also in other hybrid platforms including superconduting and trapped-ion systems, which allows us to find various efficient routes towards fault-tolerant quantum computing. 2023-12-31T10:57:31Z 21 pages, 8 figures PRX Quantum 5, 030322 (2024) Jaehak Lee Nuri Kang Seok-Hyung Lee Hyunseok Jeong Liang Jiang Seung-Woo Lee 10.1103/PRXQuantum.5.030322 http://arxiv.org/abs/2401.00469v1 Exploring the Synergy: A Review of Dual-Functional Radar Communication Systems 2023-12-31T11:55:09Z This review paper examines the concept and advancements in the evolving landscape of Dual-functional Radar Communication (DFRC) systems. Traditionally, radar and communication systems have functioned independently, but current research is actively investigating the integration of these functionalities into a unified platform. This paper discusses the motivations behind the development of DFRC systems, the challenges involved, and the potential benefits they offer. A discussion on the performance bounds for DFRC systems is also presented. The paper encompasses a comprehensive analysis of various techniques, architectures, and technologies used in the design and optimization of DFRC systems, along with their performance and trade-offs. Additionally, we explore potential application scenarios for these joint communication and sensing systems, offering a comprehensive perspective on the multifaceted landscape of DFRC technology. 2023-12-31T11:55:09Z 17 pages, 7 figures IEEE Aerospace and Electronic Systems Magazine, 41, 2026, 94-126 Ali Hanif Sajid Ahmed Tareq Y. Al-Naffouri Mohamed-Slim Alouin 10.1109/MAES.2025.3551690 http://arxiv.org/abs/2401.00405v1 Generalizing Single-View 3D Shape Retrieval to Occlusions and Unseen Objects 2023-12-31T05:39:38Z Single-view 3D shape retrieval is a challenging task that is increasingly important with the growth of available 3D data. Prior work that has studied this task has not focused on evaluating how realistic occlusions impact performance, and how shape retrieval methods generalize to scenarios where either the target 3D shape database contains unseen shapes, or the input image contains unseen objects. In this paper, we systematically evaluate single-view 3D shape retrieval along three different axes: the presence of object occlusions and truncations, generalization to unseen 3D shape data, and generalization to unseen objects in the input images. We standardize two existing datasets of real images and propose a dataset generation pipeline to produce a synthetic dataset of scenes with multiple objects exhibiting realistic occlusions. Our experiments show that training on occlusion-free data as was commonly done in prior work leads to significant performance degradation for inputs with occlusion. We find that that by first pretraining on our synthetic dataset with occlusions and then finetuning on real data, we can significantly outperform models from prior work and demonstrate robustness to both unseen 3D shapes and unseen objects. 2023-12-31T05:39:38Z Qirui Wu Daniel Ritchie Manolis Savva Angel X. Chang http://arxiv.org/abs/2401.00426v1 keqing: knowledge-based question answering is a nature chain-of-thought mentor of LLM 2023-12-31T08:39:04Z Large language models (LLMs) have exhibited remarkable performance on various natural language processing (NLP) tasks, especially for question answering. However, in the face of problems beyond the scope of knowledge, these LLMs tend to talk nonsense with a straight face, where the potential solution could be incorporating an Information Retrieval (IR) module and generating response based on these retrieved knowledge. In this paper, we present a novel framework to assist LLMs, such as ChatGPT, to retrieve question-related structured information on the knowledge graph, and demonstrate that Knowledge-based question answering (Keqing) could be a nature Chain-of-Thought (CoT) mentor to guide the LLM to sequentially find the answer entities of a complex question through interpretable logical chains. Specifically, the workflow of Keqing will execute decomposing a complex question according to predefined templates, retrieving candidate entities on knowledge graph, reasoning answers of sub-questions, and finally generating response with reasoning paths, which greatly improves the reliability of LLM's response. The experimental results on KBQA datasets show that Keqing can achieve competitive performance and illustrate the logic of answering each question. 2023-12-31T08:39:04Z 12 pages, 6 figures Chaojie Wang Yishi Xu Zhong Peng Chenxi Zhang Bo Chen Xinrun Wang Lei Feng Bo An http://arxiv.org/abs/2401.00437v1 BatchEval: Towards Human-like Text Evaluation 2023-12-31T09:34:51Z Significant progress has been made in automatic text evaluation with the introduction of large language models (LLMs) as evaluators. However, current sample-wise evaluation paradigm suffers from the following issues: (1) Sensitive to prompt design; (2) Poor resistance to noise; (3) Inferior ensemble performance with static reference. Inspired by the fact that humans treat both criterion definition and inter sample comparison as references for evaluation, we propose BatchEval, a paradigm that conducts batch-wise evaluation iteratively to alleviate the above problems. We explore variants under this paradigm and confirm the optimal settings are two stage procedure with heterogeneous batch composition strategy and decimal scoring format. Comprehensive experiments across 3 LLMs on 4 text evaluation tasks demonstrate that BatchEval outperforms state-of-the-art methods by 10.5% on Pearson correlations with only 64% API cost on average. Further analyses have been conducted to verify the robustness, generalization, and working mechanism of BatchEval. 2023-12-31T09:34:51Z 19 pages, 9 figures Peiwen Yuan Shaoxiong Feng Yiwei Li Xinglin Wang Boyuan Pan Heda Wang Kan Li http://arxiv.org/abs/2401.00440v2 TSGAN: An Optical-to-SAR Dual Conditional GAN for Optical based SAR Temporal Shifting 2024-01-04T09:43:33Z In contrast to the well-investigated field of SAR-to-Optical translation, this study explores the lesser-investigated domain of Optical-to-SAR translation, a challenging field due to the ill-posed nature of this translation. The complexity arises as a single optical data can have multiple SAR representations based on the SAR viewing geometry. We propose a novel approach, termed SAR Temporal Shifting, which inputs an optical data from the desired timestamp along with a SAR data from a different temporal point but with a consistent viewing geometry as the expected SAR data, both complemented with a change map of optical data during the intervening period. This model modifies the SAR data based on the changes observed in optical data to generate the SAR data for the desired timestamp. Our model, a dual conditional Generative Adversarial Network (GAN), named Temporal Shifting GAN (TSGAN), incorporates a siamese encoder in both the Generator and the Discriminator. To prevent the model from overfitting on the input SAR data, we employed a change weighted loss function. Our approach surpasses traditional translation methods by eliminating the GAN's fiction phenomenon, particularly in unchanged regions, resulting in higher SSIM and PSNR in these areas. Additionally, modifications to the Pix2Pix architecture and the inclusion of attention mechanisms have enhanced the model's performance on all regions of the data. This research paves the way for leveraging legacy optical datasets, the most abundant and longstanding source of Earth imagery data, extending their use to SAR domains and temporal analyses. To foster further research, we provide the code, datasets used in our study, and a framework for generating paired SAR-Optical datasets for new regions of interest. These resources are available on github.com/moienr/TemporalGAN 2023-12-31T09:38:53Z Comments: Added acknowledgments and corrected a typo. No changes to the main content Moien Rangzan Sara Attarchi Richard Gloaguen Seyed Kazem Alavipanah http://arxiv.org/abs/2401.00441v1 Quantitative unique continuation for real-valued solutions to second order elliptic equations in the plane 2023-12-31T09:45:09Z In this article, we study a quantitative form of the Landis conjecture on exponential decay for real-valued solutions to second order elliptic equations with variable coefficients in the plane. In particular, we prove the following qualitative form of Landis conjecture, for $W_1, W_2 \in L^{\infty}(\mathbb R^2;\mathbb R^2)$, $V \in L^{\infty}(\mathbb R^2;\mathbb R)$ and $u \in H_{\mathrm{loc}}^{1}(\mathbb R^2)$ a real-valued weak solution to $-Δu - \nabla \cdot ( W_1 u ) +W_2 \cdot \nabla u + V u = 0$ in $\mathbb R^2$, satisfying for $δ>0$, $|u(x)| \leq \exp(- |x|^{1+δ})$, $x \in \mathbb R^2$, then $u \equiv 0$. Our methodology of proof is inspired by the one recently developed by Logunov, Malinnikova, Nadirashvili, and Nazarov that have treated the equation $-Δu + V u = 0$ in $\mathbb R^2$. Nevertheless, several differences and additional difficulties appear. New weak quantitative maximum principles are established for the construction of a positive multiplier in a suitable perforated domain, depending on the nodal set of $u$. The resulted divergence elliptic equation is then transformed into a non-homogeneous $\partial_{\overline{z}}$ equation thanks to a generalization of Stoilow factorization theorem obtained by the theory of quasiconformal mappings, an approximate type Poincaré lemma and the use of the Cauchy transform. Finally, a suitable Carleman estimate applied to the operator $\partial_{\overline{z}}$ is the last ingredient of our proof. 2023-12-31T09:45:09Z Comments welcome Kévin Le Balc'h Diego A. Souza http://arxiv.org/abs/2401.00466v1 Online Symbolic Music Alignment with Offline Reinforcement Learning 2023-12-31T11:42:42Z Symbolic Music Alignment is the process of matching performed MIDI notes to corresponding score notes. In this paper, we introduce a reinforcement learning (RL)-based online symbolic music alignment technique. The RL agent - an attention-based neural network - iteratively estimates the current score position from local score and performance contexts. For this symbolic alignment task, environment states can be sampled exhaustively and the reward is dense, rendering a formulation as a simplified offline RL problem straightforward. We evaluate the trained agent in three ways. First, in its capacity to identify correct score positions for sampled test contexts; second, as the core technique of a complete algorithm for symbolic online note-wise alignment; and finally, as a real-time symbolic score follower. We further investigate the pitch-based score and performance representations used as the agent's inputs. To this end, we develop a second model, a two-step Dynamic Time Warping (DTW)-based offline alignment algorithm leveraging the same input representation. The proposed model outperforms a state-of-the-art reference model of offline symbolic music alignment. 2023-12-31T11:42:42Z Proceedings of the 24th International Society for Music Information Retrieval Conference, {ISMIR} 2023, Milan, Italy, November 5-9, 2023 Silvan David Peter 10.5281/zenodo.10265367 http://arxiv.org/abs/2401.00471v1 Sounding Out Reconstruction Error-Based Evaluation of Generative Models of Expressive Performance 2023-12-31T11:59:20Z Generative models of expressive piano performance are usually assessed by comparing their predictions to a reference human performance. A generative algorithm is taken to be better than competing ones if it produces performances that are closer to a human reference performance. However, expert human performers can (and do) interpret music in different ways, making for different possible references, and quantitative closeness is not necessarily aligned with perceptual similarity, raising concerns about the validity of this evaluation approach. In this work, we present a number of experiments that shed light on this problem. Using precisely measured high-quality performances of classical piano music, we carry out a listening test indicating that listeners can sometimes perceive subtle performance difference that go unnoticed under quantitative evaluation. We further present tests that indicate that such evaluation frameworks show a lot of variability in reliability and validity across different reference performances and pieces. We discuss these results and their implications for quantitative evaluation, and hope to foster a critical appreciation of the uncertainties involved in quantitative assessments of such performances within the wider music information retrieval (MIR) community. 2023-12-31T11:59:20Z 10th International Conference on Digital Libraries for Musicology, November 10, 2023, Milan, Italy Silvan David Peter Carlos Eduardo Cancino-Chacón Emmanouil Karystinaios Gerhard Widmer 10.1145/3625135.3625141 http://arxiv.org/abs/2401.00479v1 $L^p$ Maximal regularity for vector-valued Schrödinger operators 2023-12-31T12:45:48Z In this paper we consider the vector-valued Schrödinger operator $-Δ+ V$, where the potential term $V$ is a matrix-valued function whose entries belong to $L^1_{\rm loc}(\mathbb{R}^d)$ and, for every $x\in\mathbb{R}^d$, $V(x)$ is a symmetric and nonnegative definite matrix, with non positive off-diagonal terms and with eigenvalues comparable each other. For this class of potential terms we obtain maximal inequality in $L^1(\mathbb{R}^d,\mathbb{R}^m).$ Assuming further that the minimal eigenvalue of $V$ belongs to some reverse Hölder class of order $q\in(1,\infty)\cup\{\infty\}$, we obtain maximal inequality in $L^p(\mathbb{R}^d,\mathbb{R}^m)$, for $p$ in between $1$ and some $q$. 2023-12-31T12:45:48Z Davide Addona Vincenzo Leone Luca Lorenzi Abdelaziz Rhandi http://arxiv.org/abs/2401.00487v1 Spinterface Mediated Magnetic Properties of Co20Fe60B20/Alq3 Heterostructures 2023-12-31T13:01:50Z Organic semiconductors (OSCs) are suitable materials for spintronics applications as they form a spinterface when placed next to a ferromagnet, which in turn leads to novel functionalities. The evolution of spinterface can tune the global magnetic anisotropy, magnetization reversal, magnetization dynamics, etc. Planar tris-(8-hydroxyquinoline)aluminum (Alq3) OSC has shown tremendous potential for spintronics applications, thanks to its efficient spin-polarized current transport ability. Here, we establish the spinterface when the Alq3 molecules are deposited on amorphous ferromagnet Co20Fe60B20(CFB). The $π$-d hybridization in CFB/Alq3 enhances the coercive field and significantly modifies the shape and size of the magnetic domains. A $\sim$100% increase in uniaxial anisotropic energies and a reduction in magnetic damping are also evident owing to the strong interfacial hybridization. 2023-12-31T13:01:50Z Swayang Priya Mahanta Antarjami Sahoo Sagarika Nayak T. P. A. Hase Del Atkinson Subhankar Bedanta http://arxiv.org/abs/2401.00488v1 The Sonified Hertzsprung-Russell Diagram 2023-12-31T13:04:42Z Understanding the physical properties of stars, and putting these properties into the context of stellar evolution, is a core challenge in astronomical research. A key visualization in studying stellar evolution is the Hertzsprung-Russell diagram (HRD), organizing data about stellar luminosity and colour into a form that is informative about stellar structure and evolution. However, connecting the HRD with other sources of information, including stellar time series, is an outstanding challenge. Here we present a new method to turn stellar time series into sound. This method encodes physically meaningful features such that auditory comparisons between sonifications of different stars preserve astrophysical differences between them. We present an interactive multimedia version of the HRD that combines both visual and auditory components and that allows exploration of different types of stars both on and off the main sequence through both visual and auditory media. 2023-12-31T13:04:42Z 8 pages, 5 figures; accepted for publication in the proceedings of "The 28th International Conference on Auditory Display (ICAD 2023) - Special Session on Astronomical Data Sonification" Daniela Huppenkothen Juan Pampin James R. A. Davenport James Wenlock http://arxiv.org/abs/2401.00449v1 Teaching Digital Accessibility to Industry Professionals using the Community of Practice Framework: An Experience Report 2023-12-31T10:55:26Z Despite recent initiatives aimed at improving accessibility, the field of digital accessibility remains markedly behind contemporary advancements in the software industry as a large number of real world software and web applications continue to fall short of accessibility requirements. A persisting skills deficit within the existing technology workforce has been an enduring impediment, hindering organizations from delivering truly accessible software products. This, in turn, elevates the risk of isolating and excluding a substantial portion of potential users. In this paper, we report lessons learned from a training program for teaching digital accessibility using the Communities of Practice (CoP) framework to industry professionals. We recruited 66 participants from a large multi-national software company and assigned them to two groups: one participating in a CoP and the other using self-paced learning. We report experiences from designing the training program, conducting the actual training, and assessing the efficiency of the two approaches. Based on these findings, we provide recommendations for practitioners in Learning and Development teams and educators in designing accessibility courses for industry professionals. 2023-12-31T10:55:26Z To be published in International Conference on Software Engineering (ICSE'24), Software Engineering Education and Training Track Parthasarathy PD Swaroop Joshi 10.1145/3639474.3640083 http://arxiv.org/abs/2401.00443v2 Data-driven Energy Efficiency Modelling in Large-scale Networks: An Expert Knowledge and ML-based Approach 2024-06-04T10:01:55Z The energy consumption of mobile networks poses a critical challenge. Mitigating this concern necessitates the deployment and optimization of network energy-saving solutions, such as carrier shutdown, to dynamically manage network resources. Traditional optimization approaches encounter complexity due to factors like the large number of cells, stochastic traffic, channel variations, and intricate trade-offs. This paper introduces the simulated reality of communication networks (SRCON) framework, a novel, data-driven modeling paradigm that harnesses live network data and employs a blend of machine learning (ML)- and expert-based models. These mix of models accurately characterizes the functioning of network components, and predicts network energy efficiency and user equipment (UE) quality of service for any energy carrier shutdown configuration in a specific network. Distinguishing itself from existing methods, SRCON eliminates the reliance on expensive expert knowledge, drive testing, or incomplete maps for predicting network performance. This paper details the pipeline employed by SRCON to decompose the large network energy efficiency modeling problem into ML and expert-based submodels. It demonstrates how, by embracing stochasticity, and carefully crafting the relationship between such submodels, the overall computational complexity can be reduced and prediction accuracy enhanced. Results derived from real network data underscore the paradigm shift introduced by SRCON, showcasing significant gains over a state-of-the art method used by a operator for network energy efficiency modeling. The reliability of this local, data-driven modeling of the network proves to be a key asset for network energy-saving optimization. 2023-12-31T10:03:08Z 24 pages, 13 figures, submitted to IEEE Transactions on Machine Learning in Communications and Networking David López-Pérez Antonio De Domenico Nicola Piovesan Merouane Debbah http://arxiv.org/abs/2401.00442v2 A Comprehensive Overview of Fish-Eye Camera Distortion Correction Methods 2024-05-13T15:18:57Z The fisheye camera, with its unique wide field of view and other characteristics, has found extensive applications in various fields. However, the fisheye camera suffers from significant distortion compared to pinhole cameras, resulting in distorted images of captured objects. Fish-eye camera distortion is a common issue in digital image processing, requiring effective correction techniques to enhance image quality. This review provides a comprehensive overview of various methods used for fish-eye camera distortion correction. The article explores the polynomial distortion model, which utilizes polynomial functions to model and correct radial distortions. Additionally, alternative approaches such as panorama mapping, grid mapping, direct methods, and deep learning-based methods are discussed. The review highlights the advantages, limitations, and recent advancements of each method, enabling readers to make informed decisions based on their specific needs. 2023-12-31T09:49:37Z Jian Xu De-Wei Han Kang Li Jun-Jie Li Zhao-Yuan Ma http://arxiv.org/abs/2401.00425v2 Geometric BV for twisted Courant sigma models and the BRST power finesse 2024-11-05T16:02:51Z We study twisted Courant sigma models, a class of topological field theories arising from the coupling of 3D 0-/2-form BF theory and Chern-Simons theory and containing a 4-form Wess-Zumino term. They are examples of theories featuring a nonlinearly open gauge algebra, where products of field equations appear in the commutator of gauge transformations, and they are reducible gauge systems. We determine the solution to the master equation using a technique, the BRST power finesse, that combines aspects of the AKSZ construction (which applies to the untwisted model) and the general BV-BRST formalism. This allows for a geometric interpretation of the BV coefficients in the interaction terms of the master action in terms of an induced generalised connection on a 4-form twisted (pre-)Courant algebroid, its Gualtieri torsion and the basic curvature tensor. It also produces a frame independent formulation of the model. We show, moreover, that the gauge fixed action is the sum of the classical one and a BRST commutator, as expected from a Schwarz type topological field theory. 2023-12-31T08:38:23Z 50 pages; published version with added references and minor corrections and improvements JHEP 07 (2024) 115 Athanasios Chatzistavrakidis Noriaki Ikeda Larisa Jonke 10.1007/JHEP07(2024)115 http://arxiv.org/abs/2401.00411v3 Study the structure of X(3872) from its lineshape 2025-04-09T08:15:53Z We fit the invariant mass distribution of ${X(3872)}\rightarrow{J}/ψπ^+π^-$ from LHCb using the propagator for S-wave near-threshold states in effective field theory. In this way, we can directly determine the $Z$ which measures the projection of the bound state on the compact state in ${X(3872)}$. Consequently, the structure of ${X(3872)}$ can be elucidated. Moreover, the fitting result also can describe well the data for ${X(3872)}\rightarrow{D}^{0}\overline{D}^{0*}$ from Belle experiment, which demonstrate the reliability of our fitting. The fitting indicates that $Z$ is a non-vanishing value within error, which supports that $X(3872)$ has a compact short-distant core. 2023-12-31T06:54:42Z 6 pages, 2 figures, 1 table Hongge Xu Ning Yu Zuman Zhang http://arxiv.org/abs/2401.00483v2 Interacting ground states of moiré ladders 2026-03-18T13:19:19Z Moiré materials have emerged as a rich platform for exploring strong correlation effects in low dimensions, with twisted bilayer graphene (TBG) as a paradigmatic example. To distill the essential ingredients driving moiré-induced phases, a simplified one-dimensional analog -- a two-leg ladder with spatially modulated interleg hopping and a uniform magnetic flux -- was recently introduced. This model, which we refer to as the moiré ladder, features a nearly flat lowest-energy band in a suitable parameter regime, capturing the band-flattening mechanism of TBG. We investigate the ground-state phase diagram of the moiré ladder using a combination of bosonization and density matrix renormalization group (DMRG) techniques, and systematically disentangle the respective roles of the flux and the hopping modulation. At half filling, previous numerical work identified a metal-insulator transition at finite interaction strength and an unexpected ferromagnetic ground state. Revisiting this, we show that the metal-insulator transition can be understood perturbatively within bosonization, governed by the number of Fermi points. In contrast, the ferromagnetic correlations are nonperturbative and require both flux and spatial modulation -- neither alone is sufficient. We extend our analysis to other fillings: one-quarter, three-quarters, slightly above half filling (half filling plus two electrons), and slightly below half filling (half filling minus two electrons). At moderate interactions, we observe ferromagnetism below half filling and antiferromagnetism above; at stronger interactions, ferromagnetism dominates across all studied fillings. Crucially, the analysis demonstrates that periodic interleg hopping alone does not engender new correlated phases; the magnetic flux is essential for the observed unconventional behavior. 2023-12-31T12:56:09Z Phys. Rev. B 113, 104432 (2026) Paban Kumar Patra Ranjith R. Kumar Yixuan Huang Hridis K. Pal 10.1103/8lz4-s1f9