https://arxiv.org/api/z+GeFugfesBBTOQe1SL5h3z+aK4arXiv Query: search_query=&id_list=2401.00401,2401.00402,2401.00403,2401.00404,2401.00405,2401.00406,2401.00407,2401.00408,2401.00409,2401.00410,2401.00411,2401.00412,2401.00413,2401.00414,2401.00415,2401.00416,2401.00417,2401.00418,2401.00419,2401.00420,2401.00421,2401.00422,2401.00423,2401.00424,2401.00425,2401.00426,2401.00427,2401.00428,2401.00429,2401.00430,2401.00431,2401.00432,2401.00433,2401.00434,2401.00435,2401.00436,2401.00437,2401.00438,2401.00439,2401.00440,2401.00441,2401.00442,2401.00443,2401.00444,2401.00445,2401.00446,2401.00447,2401.00448,2401.00449,2401.00450,2401.00451,2401.00452,2401.00453,2401.00454,2401.00455,2401.00456,2401.00457,2401.00458,2401.00459,2401.00460,2401.00461,2401.00462,2401.00463,2401.00464,2401.00465,2401.00466,2401.00467,2401.00468,2401.00469,2401.00470,2401.00471,2401.00472,2401.00473,2401.00474,2401.00475,2401.00476,2401.00477,2401.00478,2401.00479,2401.00480,2401.00481,2401.00482,2401.00483,2401.00484,2401.00485,2401.00486,2401.00487,2401.00488,2401.00489,2401.00490,2401.00491,2401.00492,2401.00493,2401.00494,2401.00495,2401.00496,2401.00497,2401.00498,2401.00499,2401.00500&start=0&max_results=1002026-07-02T22:41:38Z1001000http://arxiv.org/abs/2401.00402v13D Multi-system Bayesian Calibration with Energy Conservation to Study Rapidity-dependent Dynamics of Nuclear Collisions2023-12-31T05:33:57ZConsiderable information about the early-stage dynamics of heavy-ion collisions is encoded in the rapidity dependence of measurements. To leverage the large amount of experimental data, we perform a systematic analysis using three-dimensional hydrodynamic simulations of multiple collision systems -- large and small, symmetric and asymmetric. Specifically, we perform fully 3D multi-stage hydrodynamic simulations initialized by a parameterized model for rapidity-dependent energy deposition, which we calibrate on the hadron multiplicity and anisotropic flow coefficients. We utilize Bayesian inference to constrain properties of the early- and late- time dynamics of the system, and highlight the impact of enforcing global energy conservation in our 3D model.2023-12-31T05:33:57ZAndi MankolliThe JETSCAPE CollaborationAaron AngeramiThe JETSCAPE CollaborationRitu AroraThe JETSCAPE CollaborationSteffen BassThe JETSCAPE CollaborationShanshan CaoThe JETSCAPE CollaborationYi ChenThe JETSCAPE CollaborationLipei DuThe JETSCAPE CollaborationRaymond EhlersThe JETSCAPE CollaborationHannah ElfnerThe JETSCAPE CollaborationWenkai FanThe JETSCAPE CollaborationRainer J. FriesThe JETSCAPE CollaborationCharles GaleThe JETSCAPE CollaborationYayun HeThe JETSCAPE CollaborationUlrich HeinzThe JETSCAPE CollaborationBarbara JacakThe JETSCAPE CollaborationPeter JacobsThe JETSCAPE CollaborationSangyong JeonThe JETSCAPE CollaborationYi JiThe JETSCAPE CollaborationLauren KasperThe JETSCAPE CollaborationMichael KordellThe JETSCAPE CollaborationAmit KumarThe JETSCAPE CollaborationR. Kunnawalkam-ElayavalliThe JETSCAPE CollaborationJoseph LatessaThe JETSCAPE CollaborationSook H. LeeThe JETSCAPE CollaborationYen-Jie LeeThe JETSCAPE CollaborationDananjaya LiyanageThe JETSCAPE CollaborationMatt LuzumThe JETSCAPE CollaborationAbhijit MajumderThe JETSCAPE CollaborationSimon MakThe JETSCAPE CollaborationChristal MartinThe JETSCAPE CollaborationHaydar MehryarThe JETSCAPE CollaborationTanner MengelThe JETSCAPE CollaborationJames MulliganThe JETSCAPE CollaborationChristine NattrassThe JETSCAPE CollaborationJean-Francois PaquetThe JETSCAPE CollaborationCameron ParkerThe JETSCAPE CollaborationJoern H. PutschkeThe JETSCAPE CollaborationGunther RolandThe JETSCAPE CollaborationBjoern SchenkeThe JETSCAPE CollaborationLoren SchwiebertThe JETSCAPE CollaborationArjun SenguptaThe JETSCAPE CollaborationChun ShenThe JETSCAPE CollaborationChathuranga SirimannaThe JETSCAPE CollaborationRon A. SoltzThe JETSCAPE CollaborationIsmail SoudiThe JETSCAPE CollaborationMichael StricklandThe JETSCAPE CollaborationYasuki TachibanaThe JETSCAPE CollaborationJulia VelkovskaThe JETSCAPE CollaborationGojko VujanovicThe JETSCAPE CollaborationXin-Nian WangThe JETSCAPE CollaborationWenbin ZhaoThe JETSCAPE Collaborationhttp://arxiv.org/abs/2401.00418v1Bounds on the minimum distance of locally recoverable codes2023-12-31T07:47:16ZWe consider locally recoverable codes (LRCs) and aim to determine the smallest possible length $n=n_q(k,d,r)$ of a linear $[n,k,d]_q$-code with locality $r$. For $k\le 7$ we exactly determine all values of $n_2(k,d,2)$ and for $k\le 6$ we exactly determine all values of $n_2(k,d,1)$. For the ternary field we also state a few numerical results. As a general result we prove that $n_q(k,d,r)$ equals the Griesmer bound if the minimum Hamming distance $d$ is sufficiently large and all other parameters are fixed.2023-12-31T07:47:16Z23 pages, 3 tablesSascha Kurzhttp://arxiv.org/abs/2401.00420v2SynCDR : Training Cross Domain Retrieval Models with Synthetic Data2024-03-19T16:56:53ZIn cross-domain retrieval, a model is required to identify images from the same semantic category across two visual domains. For instance, given a sketch of an object, a model needs to retrieve a real image of it from an online store's catalog. A standard approach for such a problem is learning a feature space of images where Euclidean distances reflect similarity. Even without human annotations, which may be expensive to acquire, prior methods function reasonably well using unlabeled images for training. Our problem constraint takes this further to scenarios where the two domains do not necessarily share any common categories in training data. This can occur when the two domains in question come from different versions of some biometric sensor recording identities of different people. We posit a simple solution, which is to generate synthetic data to fill in these missing category examples across domains. This, we do via category preserving translation of images from one visual domain to another. We compare approaches specifically trained for this translation for a pair of domains, as well as those that can use large-scale pre-trained text-to-image diffusion models via prompts, and find that the latter can generate better replacement synthetic data, leading to more accurate cross-domain retrieval models. Our best SynCDR model can outperform prior art by up to 15\%. Code for our work is available at https://github.com/samarth4149/SynCDR .2023-12-31T08:06:53ZPre-printSamarth MishraCarlos D. CastilloHongcheng WangKate SaenkoVenkatesh Saligramahttp://arxiv.org/abs/2401.00421v1From Text to Pixels: A Context-Aware Semantic Synergy Solution for Infrared and Visible Image Fusion2023-12-31T08:13:47ZWith the rapid progression of deep learning technologies, multi-modality image fusion has become increasingly prevalent in object detection tasks. Despite its popularity, the inherent disparities in how different sources depict scene content make fusion a challenging problem. Current fusion methodologies identify shared characteristics between the two modalities and integrate them within this shared domain using either iterative optimization or deep learning architectures, which often neglect the intricate semantic relationships between modalities, resulting in a superficial understanding of inter-modal connections and, consequently, suboptimal fusion outcomes. To address this, we introduce a text-guided multi-modality image fusion method that leverages the high-level semantics from textual descriptions to integrate semantics from infrared and visible images. This method capitalizes on the complementary characteristics of diverse modalities, bolstering both the accuracy and robustness of object detection. The codebook is utilized to enhance a streamlined and concise depiction of the fused intra- and inter-domain dynamics, fine-tuned for optimal performance in detection tasks. We present a bilevel optimization strategy that establishes a nexus between the joint problem of fusion and detection, optimizing both processes concurrently. Furthermore, we introduce the first dataset of paired infrared and visible images accompanied by text prompts, paving the way for future research. Extensive experiments on several datasets demonstrate that our method not only produces visually superior fusion results but also achieves a higher detection mAP over existing methods, achieving state-of-the-art results.2023-12-31T08:13:47Z10 pages, 12 figures, 3 tables, conferenceXingyuan LiYang ZouJinyuan LiuZhiying JiangLong MaXin FanRisheng Liuhttp://arxiv.org/abs/2401.00424v1SDIF-DA: A Shallow-to-Deep Interaction Framework with Data Augmentation for Multi-modal Intent Detection2023-12-31T08:33:37ZMulti-modal intent detection aims to utilize various modalities to understand the user's intentions, which is essential for the deployment of dialogue systems in real-world scenarios. The two core challenges for multi-modal intent detection are (1) how to effectively align and fuse different features of modalities and (2) the limited labeled multi-modal intent training data. In this work, we introduce a shallow-to-deep interaction framework with data augmentation (SDIF-DA) to address the above challenges. Firstly, SDIF-DA leverages a shallow-to-deep interaction module to progressively and effectively align and fuse features across text, video, and audio modalities. Secondly, we propose a ChatGPT-based data augmentation approach to automatically augment sufficient training data. Experimental results demonstrate that SDIF-DA can effectively align and fuse multi-modal features by achieving state-of-the-art performance. In addition, extensive analyses show that the introduced data augmentation approach can successfully distill knowledge from the large language model.2023-12-31T08:33:37ZAccepted by ICASSP 2024Shijue HuangLibo QinBingbing WangGeng TuRuifeng Xuhttp://arxiv.org/abs/2401.00429v1Deeper and Wider Networks for Performance Metrics Prediction in Communication Networks2023-12-31T08:59:59ZIn today's era, users have increasingly high expectations regarding the performance and efficiency of communication networks. Network operators aspire to achieve efficient network planning, operation, and optimization through Digital Twin Networks (DTN). The effectiveness of DTN heavily relies on the network model, with graph neural networks (GNN) playing a crucial role in network modeling. However, existing network modeling methods still lack a comprehensive understanding of communication networks. In this paper, we propose DWNet (Deeper and Wider Networks), a heterogeneous graph neural network modeling method based on data-driven approaches that aims to address end-to-end latency and jitter prediction in network models. This method stands out due to two distinctive features: firstly, it introduces deeper levels of state participation in the message passing process; secondly, it extensively integrates relevant features during the feature fusion process. Through experimental validation and evaluation, our model achieves higher prediction accuracy compared to previous research achievements, particularly when dealing with unseen network topologies during model training. Our model not only provides more accurate predictions but also demonstrates stronger generalization capabilities across diverse topological structures.2023-12-31T08:59:59ZAijia LiuShiqing LiuXiaobing Peihttp://arxiv.org/abs/2401.00438v1SFGANS Self-supervised Future Generator for human ActioN Segmentation2023-12-31T09:36:55ZThe ability to locate and classify action segments in long untrimmed video is of particular interest to many applications such as autonomous cars, robotics and healthcare applications. Today, the most popular pipeline for action segmentation is composed of encoding the frames into feature vectors, which are then processed by a temporal model for segmentation. In this paper we present a self-supervised method that comes in the middle of the standard pipeline and generated refined representations of the original feature vectors. Experiments show that this method improves the performance of existing models on different sub-tasks of action segmentation, even without additional hyper parameter tuning.2023-12-31T09:36:55ZOr BermanAdam GoldbraikhShlomi Lauferhttp://arxiv.org/abs/2401.00439v1On the breathing of spectral bands in periodic quantum waveguides with inflating resonators2023-12-31T09:38:37ZWe are interested in the lower part of the spectrum of the Dirichlet Laplacian $A^\varepsilon$ in a thin waveguide $Π^\varepsilon$ obtained by repeating periodically a pattern, itself constructed by scaling an inner field geometry $Ω$ by a small factor $\varepsilon>0$. The Floquet-Bloch theory ensures that the spectrum of $A^\varepsilon$ has a band-gap structure. Due to the Dirichlet boundary conditions, these bands all move to $+\infty$ as $O(\varepsilon^{-2})$ when $\varepsilon\to0^+$. Concerning their widths, applying techniques of dimension reduction, we show that the results depend on the dimension of the so-called space of almost standing waves in $Ω$ that we denote by $\mathrm{X}_\dagger$. Generically, i.e. for most $Ω$, there holds $\mathrm{X}_\dagger=\{0\}$ and the lower part of the spectrum of $A^\varepsilon$ is very sparse, made of bands of length at most $O(\varepsilon)$ as $\varepsilon\to0^+$. For certain $Ω$ however, we have $\mathrm{dim}\,\mathrm{X}_\dagger=1$ and then there are bands of length $O(1)$ which allow for wave propagation in $Π^\varepsilon$. The main originality of this work lies in the study of the behaviour of the spectral bands when perturbing $Ω$ around a particular $Ω_\star$ where $\mathrm{dim}\,\mathrm{X}_\dagger=1$. We show a breathing phenomenon for the spectrum of $A^\varepsilon$: when inflating $Ω$ around $Ω_\star$, the spectral bands rapidly expand before shrinking. In the process, a band dives below the normalized threshold $π^2/\varepsilon^2$, stops breathing and becomes extremely short as $Ω$ continues to inflate.2023-12-31T09:38:37ZLucas ChesnelSergei A. Nazarovhttp://arxiv.org/abs/2401.00453v2Global well-posedness for the Cauchy problem of the Zakharov-Kuznetsov equation on cylindrical spaces2024-01-02T12:46:37ZWe prove that the Zakharov-Kuznetsov equation on cylindrical spaces is globally well-posed below the energy norm. As is known, local well-posedness below energy space was obtained by the first author. We adapt I-method to extend the solutions globally in time. Using modified energies, we obtain the polynomial bounds on the $H^s$ growth for the global solutions.2023-12-31T11:04:03Z24 pagesSatoshi OsawaHideo Takaokahttp://arxiv.org/abs/2401.00460v1RainSD: Rain Style Diversification Module for Image Synthesis Enhancement using Feature-Level Style Distribution2023-12-31T11:30:42ZAutonomous driving technology nowadays targets to level 4 or beyond, but the researchers are faced with some limitations for developing reliable driving algorithms in diverse challenges. To promote the autonomous vehicles to spread widely, it is important to address safety issues on this technology. Among various safety concerns, the sensor blockage problem by severe weather conditions can be one of the most frequent threats for multi-task learning based perception algorithms during autonomous driving. To handle this problem, the importance of the generation of proper datasets is becoming more significant. In this paper, a synthetic road dataset with sensor blockage generated from real road dataset BDD100K is suggested in the format of BDD100K annotation. Rain streaks for each frame were made by an experimentally established equation and translated utilizing the image-to-image translation network based on style transfer. Using this dataset, the degradation of the diverse multi-task networks for autonomous driving, such as lane detection, driving area segmentation, and traffic object detection, has been thoroughly evaluated and analyzed. The tendency of the performance degradation of deep neural network-based perception systems for autonomous vehicle has been analyzed in depth. Finally, we discuss the limitation and the future directions of the deep neural network-based perception algorithms and autonomous driving dataset generation based on image-to-image translation.2023-12-31T11:30:42ZUnder ReviewHyeonjae JeonJunghyun SeoTaesoo KimSungho SonJungki LeeGyeungho ChoiYongseob Limhttp://arxiv.org/abs/2401.00461v1A Penalized Functional Linear Cox Regression Model for Spatially-defined Environmental Exposure with an Estimated Buffer Distance2023-12-31T11:31:57ZIn environmental health research, it is of interest to understand the effect of the neighborhood environment on health. Researchers have shown a protective association between green space around a person's residential address and depression outcomes. In measuring exposure to green space, distance buffers are often used. However, buffer distances differ across studies. Typically, the buffer distance is determined by researchers a priori. It is unclear how to identify an appropriate buffer distance for exposure assessment. To address geographic uncertainty problem for exposure assessment, we present a domain selection algorithm based on the penalized functional linear Cox regression model. The theoretical properties of our proposed method are studied and simulation studies are conducted to evaluate finite sample performances of our method. The proposed method is illustrated in a study of associations of green space exposure with depression and/or antidepressant use in the Nurses' Health Study.2023-12-31T11:31:57Z27 pages, 5 figuresJooyoung LeeZhibing HeCharlotte RoscoePeter JamesLi XuDonna SpiegelmanDavid ZuckerMolin Wanghttp://arxiv.org/abs/2401.00465v1V2X communication coverage analysis for connected vehicles in intelligent transportation networks: A case study for the city of Xanthi, Greece2023-12-31T11:39:18ZIntelligent transportation systems (ITS) have been developed to improve traffic flow, efficiency, and safety in transportation. Technological advancements in communication such as the Vehicle-to-Everything (V2X), Vehicle-to-Vehicle (V2V) and Vehicle-to Infrastructure (V2I) enable the real-time exchange of information between vehicles and other entities on the road network, and thus play a significant role in their safety and efficiency. This paper presents a simulation study that models V2V and V2I communication to identify the most suitable range of data transmission between vehicles and infrastructure. The provincial city of Xanthi, Greece is used as a cases study, and the goal is to evaluate whether the proposed placement of Road Side Unit (RSU) provided adequate communication coverage on the city's road network. An analysis through different scenarios identified improvements in traffic management, driving behavior and environmental conditions under different RSU coverage. The results highlight that the communication range of 400 meters is the most adequate option for optimum traffic management in the city of Xanthi.2023-12-31T11:39:18ZWireless World Research Forum, Meeting 49, March 28th-30th 2023, Poznań, Poland, Towards sustainable and automated communicationsEvangelos BazinasAndreas GregoriadesMarios RaspopoulosMichael Georgiadeshttp://arxiv.org/abs/2401.00478v1Partial classification of the large-time behavior of solutions to cubic nonlinear Schrödinger systems2023-12-31T12:44:13ZIn this paper, we study the large-time behavior of small solutions to the standard form of the systems of 1D cubic nonlinear Schrödinger equations consisting of two components and possessing a coercive mass-like conserved quantity. The cubic nonlinearity is known to be critical in one space dimension in view of the large-time behavior. By employing the result by Katayama and Sakoda, one can obtain the large-time behavior of the solution if we can integrate the corresponding ODE system. We introduce an integration scheme suited to the system. The key idea is to rewrite the ODE system, which is cubic, as a quadratic system of quadratic quantities of the original unknown. By using this technique, we described the large-time behavior of solutions in terms of elementary functions and the Jacobi elliptic functions for several examples of standard systems.2023-12-31T12:44:13Z47 pages, no figureSatoshi Masakihttp://arxiv.org/abs/2401.00481v3Analysis of the $\mathrm{X_{AV}}$ state through its electromagnetic properties2024-03-26T10:44:18ZTo improve our understanding of the quark-gluon dynamics underlying multiquark states, we systematically study their electromagnetic properties. In this study, the magnetic and quadrupole moments of the theoretically predicted singly-charmed state with the quantum numbers $\mathrm{J^P = 1^+}$ is investigated within the framework of the QCD light-cone sum rules method by considering the diquark-antidiquark configuration of this state with quark contents $[ud][\bar{c}\bar{s}]$. The predicted results for the magnetic and quadrupole moments are as $μ_{\mathrm{X_{AV}}}=-0.89 ^{+0.14}_{-0.12}~μ_N $ and $\mathcal{D}_{\mathrm{X_{AV}}} = (-0.46 ^{+0.07}_{-0.06})\times 10^{-2} ~\mbox{fm}^2$. The results obtained can be useful in determining the exact nature of this state. This work will hopefully stimulate experimental interest in the study of the electromagnetic properties of multiquark systems.2023-12-31T12:46:42Z8 pages, 1 tables, 1 figure, version accepted by European Physical Journal CU. Özdemhttp://arxiv.org/abs/2401.00493v1Reduced variance random batch methods for nonlocal PDEs2023-12-31T13:27:02ZRandom Batch Methods (RBM) for mean-field interacting particle systems enable the reduction of the quadratic computational cost associated with particle interactions to a near-linear cost. The essence of these algorithms lies in the random partitioning of the particle ensemble into smaller batches at each time step. The interaction of each particle within these batches is then evolved until the subsequent time step. This approach effectively decreases the computational cost by an order of magnitude while increasing the amount of fluctuations due to the random partitioning. In this work, we propose a variance reduction technique for RBM applied to nonlocal PDEs of Fokker-Planck type based on a control variate strategy. The core idea is to construct a surrogate model that can be computed on the full set of particles at a linear cost while maintaining enough correlations with the original particle dynamics. Examples from models of collective behavior in opinion spreading and swarming dynamics demonstrate the great potential of the present approach.2023-12-31T13:27:02ZLorenzo PareschiMattia Zanellahttp://arxiv.org/abs/2401.00459v3Interfacial dripping faucet: generating monodisperse liquid lenses2024-12-17T16:44:10ZWe present a surface analog to a dripping faucet, where a viscous liquid slides down an immiscible meniscus. Periodic pinch-off of the dripping filament is observed, generating a succession of monodisperse floating lenses. We show that this interfacial dripping faucet can be described analogously to its single-phase counterpart, replacing surface tension by the spreading coefficient, and even undergoes a transition to a jetting regime. This liquid/liquid/gas system opens perspectives for the study of the dynamics of emulsions at interfaces.2023-12-31T11:26:13Z13 pages, 10 figures, 5 supplemental moviesLorène ChampougnyVincent BertinJacco H. SnoeijerJavier Rodríguez-Rodríguez10.1103/PhysRevLett.133.254001http://arxiv.org/abs/2401.00475v3E-chat: Emotion-sensitive Spoken Dialogue System with Large Language Models2024-07-27T07:45:43ZThis study focuses on emotion-sensitive spoken dialogue in human-machine speech interaction. With the advancement of Large Language Models (LLMs), dialogue systems can handle multimodal data, including audio. Recent models have enhanced the understanding of complex audio signals through the integration of various audio events. However, they are unable to generate appropriate responses based on emotional speech. To address this, we introduce the Emotional chat Model (E-chat), a novel spoken dialogue system capable of comprehending and responding to emotions conveyed from speech. This model leverages an emotion embedding extracted by a speech encoder, combined with LLMs, enabling it to respond according to different emotional contexts. Additionally, we introduce the E-chat200 dataset, designed explicitly for emotion-sensitive spoken dialogue. In various evaluation metrics, E-chat consistently outperforms baseline model, demonstrating its potential in emotional comprehension and human-machine interaction.2023-12-31T12:29:12Z5 pages, 3 figuresHongfei XueYuhao LiangBingshen MuShiliang ZhangMengzhe ChenQian ChenLei Xiehttp://arxiv.org/abs/2401.00456v2Double-well Net for Image Segmentation2024-07-28T08:40:34ZIn this study, our goal is to integrate classical mathematical models with deep neural networks by introducing two novel deep neural network models for image segmentation known as Double-well Nets. Drawing inspirations from the Potts model, our models leverage neural networks to represent a region force functional. We extend the well-know MBO (Merriman-Bence-Osher) scheme to solve the Potts model. The widely recognized Potts model is approximated using a double-well potential and then solved by an operator-splitting method, which turns out to be an extension of the well-known MBO scheme. Subsequently, we replace the region force functional in the Potts model with a UNet-type network, which is data-driven and is designed to capture multiscale features of images, and also introduce control variables to enhance effectiveness. The resulting algorithm is a neural network activated by a function that minimizes the double-well potential. What sets our proposed Double-well Nets apart from many existing deep learning methods for image segmentation is their strong mathematical foundation. They are derived from the network approximation theory and employ the MBO scheme to approximately solve the Potts model. By incorporating mathematical principles, Double-well Nets bridge the MBO scheme and neural networks, and offer an alternative perspective for designing networks with mathematical backgrounds. Through comprehensive experiments, we demonstrate the performance of Double-well Nets, showcasing their superior accuracy and robustness compared to state-of-the-art neural networks. Overall, our work represents a valuable contribution to the field of image segmentation by combining the strengths of classical variational models and deep neural networks. The Double-well Nets introduce an innovative approach that leverages mathematical foundations to enhance segmentation performance.2023-12-31T11:16:12ZHao LiuJun LiuRaymond H. ChanXue-Cheng Taihttp://arxiv.org/abs/2401.00451v1Exploring the Need of Accessibility Education in the Software Industry: Insights from a Survey of Software Professionals in India2023-12-31T10:58:30ZA UserWay study in 2021 indicates that an annual global e-commerce revenue loss of approximately $16 billion can be attributed to inaccessible websites and applications. According to the 2023 WebAIM study, only 3.7% of the world's top one million website homepages are fully accessible. This shows that many software developers use poor coding practices that don't adhere to the Web Content Accessibility Guidelines (WCAG). This research centers on software professionals and their role in addressing accessibility. This work seeks to understand (a) who within the software development community actively practices accessibility, (b) when and how accessibility is considered in the software development lifecycle, (c) the various challenges encountered in building accessible software, and (d) the resources required by software professionals to enhance product accessibility. Our survey of 269 software professionals from India sheds light on the pressing need for accessibility education within the software industry. A substantial majority (69.9%, N=269) of respondents express the need for training materials, workshops, and bootcamps to enhance their accessibility skills. We present a list of actionable recommendations that can be implemented within the industry to promote accessibility awareness and skills. We also open source our raw data for further research, encouraging continued exploration in this domain.2023-12-31T10:58:30ZTo be published in International Conference on Software Engineering (ICSE'24), Software Engineering Education and Training TrackParthasarathy P DSwaroop Joshi10.1145/3639474.3640079http://arxiv.org/abs/2401.00433v4Quantized collision invariants on the sphere2024-04-20T18:29:05ZWe show that a measurable function $g:\mathbb{S}^{d-1}\to\mathbb{R}$, with $d\geq 3$, satisfies the functional relation \begin{equation*} g(ω)+g(ω_*)=g(ω')+g(ω_*'), \end{equation*} for all admissible $ω,ω_*,ω',ω_*'\in\mathbb{S}^{d-1}$ in the sense that \begin{equation*} ω+ω_*=ω'+ω_*', \end{equation*} if and only if it can be written as \begin{equation*} g(ω)=A+B\cdotω, \end{equation*} for some constants $A\in \mathbb{R}$ and $B\in\mathbb{R}^d$.
Such functions form a family of quantized collision invariants which play a fundamental role in the study of hydrodynamic regimes of the Boltzmann--Fermi--Dirac equation near Fermionic condensates, i.e., at low temperatures. In particular, they characterize the elastic collisional dynamics of Fermions near a statistical equilibrium where quantum effects are predominant.2023-12-31T09:15:00ZCommunications in Mathematics, Volume 32 (2024), Issue 3 (Special issue: Portuguese Mathematics) (April 25, 2024) cm:12766Benjamin AnwasiaDiogo Arsénio10.46298/cm.12766http://arxiv.org/abs/2401.00497v2The vector-valued Stieltjes moment problem with general exponents2024-06-24T07:43:37ZWe characterize the sequences of complex numbers $(z_{n})_{n \in \mathbb{N}}$ and the locally complete $(DF)$-spaces $E$ such that for each $(e_{n})_{n \in \mathbb{N}} \in E^\mathbb{N}$ there exists an $E$-valued function $\mathbf{f}$ on $(0,\infty)$ (satisfying a mild regularity condition) such that $$\int_{0}^{\infty} t^{z_{n}} \mathbf{f}(t) dt = e_{n}, \qquad \forall n \in \mathbb{N},$$ where the integral should be understood as a Pettis integral. Moreover, in this case, we show that there always exists a solution $\mathbf{f}$ that is smooth on $(0,\infty)$ and satisfies certain optimal growth bounds near $0$ and $\infty$. The scalar-valued case $(E = \mathbb{C})$ was treated by Durán [Math. Nachr. 158 (1992), 175-194]. Our work is based upon his result.2023-12-31T13:35:11Z13 pagesAndreas DebrouwereLenny Neyt10.1007/s43037-024-00364-8http://arxiv.org/abs/2401.00448v3Beyond Chinchilla-Optimal: Accounting for Inference in Language Model Scaling Laws2025-04-14T10:11:13ZLarge language model (LLM) scaling laws are empirical formulas that estimate changes in model quality as a result of increasing parameter count and training data. However, these formulas, including the popular Deepmind Chinchilla scaling laws, neglect to include the cost of inference. We modify the Chinchilla scaling laws to calculate the optimal LLM parameter count and pre-training data size to train and deploy a model of a given quality and inference demand. We conduct our analysis both in terms of a compute budget and real-world costs and find that LLM researchers expecting reasonably large inference demand (~1B requests) should train models smaller and longer than Chinchilla-optimal. Furthermore, we train 47 models of varying sizes and parameter counts to validate our formula and find that model quality continues to improve as we scale tokens per parameter to extreme ranges (up to 10,000). Finally, we ablate the procedure used to fit the Chinchilla scaling law coefficients and find that developing scaling laws only from data collected at typical token/parameter ratios overestimates the impact of additional tokens at these extreme ranges.2023-12-31T10:53:58Z16 pages, 7 figures, In the 41st International Conference on Machine Learning, 2024Nikhil SardanaJacob PortesSasha DoubovJonathan Franklehttp://arxiv.org/abs/2401.00431v2Wild2Avatar: Rendering Humans Behind Occlusions2025-08-14T22:41:09ZRendering the visual appearance of moving humans from occluded monocular videos is a challenging task. Most existing research renders 3D humans under ideal conditions, requiring a clear and unobstructed scene. Those methods cannot be used to render humans in real-world scenes where obstacles may block the camera's view and lead to partial occlusions. In this work, we present Wild2Avatar, a neural rendering approach catered for occluded in-the-wild monocular videos. We propose occlusion-aware scene parameterization for decoupling the scene into three parts - occlusion, human, and background. Additionally, extensive objective functions are designed to help enforce the decoupling of the human from both the occlusion and the background and to ensure the completeness of the human model. We verify the effectiveness of our approach with experiments on in-the-wild videos.2023-12-31T09:01:34ZIEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI). Webpage: https://cs.stanford.edu/~xtiange/projects/wild2avatar/Tiange XiangAdam SunScott DelpKazuki KozukaLi Fei-FeiEhsan Adelihttp://arxiv.org/abs/2401.00422v3Interpreting the Curse of Dimensionality from Distance Concentration and Manifold Effect2025-03-20T10:08:31ZThe characteristics of data like distribution and heterogeneity, become more complex and counterintuitive as dimensionality increases. This phenomenon is known as curse of dimensionality, where common patterns and relationships (e.g., internal pattern and boundary pattern) that hold in low-dimensional space may be invalid in higher-dimensional space. It leads to a decreasing performance for the regression, classification, or clustering models or algorithms. Curse of dimensionality can be attributed to many causes. In this paper, we first summarize the potential challenges associated with manipulating high-dimensional data, and explains the possible causes for the failure of regression, classification, or clustering tasks. Subsequently, we delve into two major causes of the curse of dimensionality, distance concentration, and manifold effect, by performing theoretical and empirical analyses. The results demonstrate that, as the dimensionality increases, nearest neighbor search (NNS) using three classical distance measurements, Minkowski distance, Chebyshev distance, and cosine distance, becomes meaningless. Meanwhile, the data incorporates more redundant features, and the variance contribution of principal component analysis (PCA) is skewed towards a few dimensions.2023-12-31T08:22:51Z21 pages, 10 figuresDehua PengZhipeng GuiHuayi Wuhttp://arxiv.org/abs/2401.00410v1Electrical and thermal transport properties of kagome metals AV$_3$Sb$_5$ (A=K, Rb, Cs)2023-12-31T06:48:14ZThe interplay between lattice geometry, band topology and electronic correlations in the newly discovered kagome compounds AV$_3$Sb$_5$ (A=K, Rb, Cs) makes this family a novel playground to investigate emergent quantum phenomena, such as unconventional superconductivity, chiral charge density wave and electronic nematicity. These exotic quantum phases naturally leave nontrivial fingerprints in transport properties of AV$_3$Sb$_5$, both in electrical and thermal channels, which are prominent probes to uncover the underlying mechanisms. In this brief review, we highlight the unusual electrical and thermal transport properties observed in the unconventional charge ordered state of AV3Sb5, including giant anomalous Hall, anomalous Nernst, ambipolar Nernst and anomalous thermal Hall effects. Connections of these anomalous transport properties to time-reversal symmetry breaking, topological and multiband fermiology, as well as electronic nematicity, are also discussed. Finally, a perspective together with challenges of this rapid growing field are given.2023-12-31T06:48:14Z34 pages,9 figures,an review article published in Tungsten 5,300(2023)Tungsten 5,300(2023)Xinrun MiKunya YangYuhan GanLong ZhangAifeng WangYisheng ChaiXiaoyuan ZhouMingquan He10.1007/s42864-022-00192-zhttp://arxiv.org/abs/2401.00492v2Edge statistics for random band matrices2025-06-03T07:35:20ZWe consider Hermitian and symmetric random band matrices on the $d$-dimensional lattice $(\mathbb{Z}/L\mathbb{Z})^d$ with bandwidth $W$, focusing on local eigenvalue statistics at the spectral edge in the limit $W\to\infty$. Our analysis reveals a critical dimension $d_c=6$ and identifies the critical bandwidth scaling as $W_c=L^{(1-d/6)_+}$. In the Hermitian case, we establish the Anderson transition for all dimensions $d<4$, and GUE edge universality when $d\geq 4$ under the condition $W\geq L^{1/3+ε}$ for any $ε>0$. In the symmetric case, we also establish parallel but more subtle transition phenomena after tadpole diagram renormalization. These findings extend Sodin's pioneering work [Ann. Math. 172, 2010], which was limited to the one-dimensional case and did not address the critical phenomena.2023-12-31T13:25:37ZPage 88, figures 14, Added Section 5 on tadpole diagram renormalization enables our results as better for symmetric case to as for Hermitian; We remove power-law band matrices that will be treated in a separate paperDang-Zheng LiuGuangyi Zouhttp://arxiv.org/abs/2401.00415v1A Novel Estimation Method for Temperature of Magnetic Nanoparticles Dominated by Brownian Relaxation Based on Magnetic Particle Spectroscopy2023-12-31T07:29:31ZThis paper presents a novel method for estimating the temperature of magnetic nanoparticles (MNPs) based on AC magnetization harmonics of MNPs dominated by Brownian relaxation. The difference in the AC magnetization response and magnetization harmonic between the Fokker-Planck equation and the Langevin function was analyzed, and we studied the relationship between the magnetization harmonic and the key factors, such as Brownian relaxation time, temperature, magnetic field strength, core size and hydrodynamic size of MNPs, excitation frequency, and so on. We proposed a compensation function for AC magnetization harmonic with consideration of the key factors and the difference between the Fokker-Planck equation and the Langevin function. Then a temperature estimation model based on the compensation function and the Langevin function was established. By employing the least squares algorithm, the temperature was successfully calculated. The experimental results show that the temperature error is less than 0.035 K in the temperature range from 310 K to 320 K. The temperature estimation model is expected to improve the performance of the magnetic nanoparticle thermometer and be applied to magnetic nanoparticle-mediated hyperthermia.2023-12-31T07:29:31ZZhongzhou DuGaoli ZhaoZhanpeng HuaNa YeYi SunWenjie WuHaochen ZhangLongtu YuShijie HanHaozhe WangWenzhong LiuTakashi Yoshidahttp://arxiv.org/abs/2401.00419v1Approximation algorithms for Job Scheduling with reconfigurable resources2023-12-31T07:50:58ZWe consider here the MultiBot problem for the scheduling and the resource parametrization of jobs related to the production or the transportation of different products inside a given time horizon. Those jobs must meet known in advance demands. The time horizon is divided into several discrete identical periods representing each the time needed to proceed a job. The objective is to find a parametrization and a schedule for the jobs in such a way they require as less resources as possible. Though this problem derived from the applicative context of reconfigurable robots, we focus here on fundamental issues. We show that the resulting strongly NP-hard Multibot problem may be handled in a greedy way with an approximation ratio of $\frac{4}{3}$.2023-12-31T07:50:58Z26 pages, 4 figuresPierre BergéMari ChaikovskaiaJean-Philippe GayonAlain Quilliothttp://arxiv.org/abs/2401.00430v2Brain-Conditional Multimodal Synthesis: A Survey and Taxonomy2024-01-03T08:50:27ZIn the era of Artificial Intelligence Generated Content (AIGC), conditional multimodal synthesis technologies (e.g., text-to-image, text-to-video, text-to-audio, etc) are gradually reshaping the natural content in the real world. The key to multimodal synthesis technology is to establish the mapping relationship between different modalities. Brain signals, serving as potential reflections of how the brain interprets external information, exhibit a distinctive One-to-Many correspondence with various external modalities. This correspondence makes brain signals emerge as a promising guiding condition for multimodal content synthesis. Brian-conditional multimodal synthesis refers to decoding brain signals back to perceptual experience, which is crucial for developing practical brain-computer interface systems and unraveling complex mechanisms underlying how the brain perceives and comprehends external stimuli. This survey comprehensively examines the emerging field of AIGC-based Brain-conditional Multimodal Synthesis, termed AIGC-Brain, to delineate the current landscape and future directions. To begin, related brain neuroimaging datasets, functional brain regions, and mainstream generative models are introduced as the foundation of AIGC-Brain decoding and analysis. Next, we provide a comprehensive taxonomy for AIGC-Brain decoding models and present task-specific representative work and detailed implementation strategies to facilitate comparison and in-depth analysis. Quality assessments are then introduced for both qualitative and quantitative evaluation. Finally, this survey explores insights gained, providing current challenges and outlining prospects of AIGC-Brain. Being the inaugural survey in this domain, this paper paves the way for the progress of AIGC-Brain research, offering a foundational overview to guide future work.2023-12-31T09:00:40ZWeijian MaiJian ZhangPengfei FangZhijun Zhanghttp://arxiv.org/abs/2401.00432v1Magnetic order and strongly-correlated effects in the one-dimensional Ising-Kondo lattice2023-12-31T09:04:36ZWe investigate the magnetic order and related strongly-correlated effects in an one-dimensional Ising-Kondo lattice with transverse field. This model is the anisotropic limit of the conventional isotropic Kondo lattice model, in the sense that the itinerant electrons interact with the localized magnetic moments via only longitudinal Kondo exchange. Adopting the numerical density-matrix-renormalization group method, we map out the ground-state phase diagram in various parameter spaces. Depending on the Kondo coupling and filling number, three distinct phases, including a metallic paramagnetic, a metallic ferromagnetic, and a gapped spin-density wave phase, are obtained. The spin-density wave is characterized by an ordering wave vector which coincides with the nesting wave vector of the Fermi surface. This makes the corresponding magnetic transition a spin analog of the Peierls transition occurring in the one-dimensional metal. Moreover, by analyzing the momentum distribution function and charge correlation function, the conduction electrons are shown to behave like free spinless fermions in the ferromagnetic phase. We finally discuss the effect of the repulsive Hubbard interaction between conduction electrons. Our work enriches the Kondo physics and deepens the current understanding of the heavy fermion compounds.2023-12-31T09:04:36Z9 pages, 12 figuresXiaofan ZhouJingtao FanSuotang Jiahttp://arxiv.org/abs/2401.00434v2GeoGalactica: A Scientific Large Language Model in Geoscience2024-04-13T17:05:03ZLarge language models (LLMs) have achieved huge success for their general knowledge and ability to solve a wide spectrum of tasks in natural language processing (NLP). Due to their impressive abilities, LLMs have shed light on potential inter-discipline applications to foster scientific discoveries of a specific domain by using artificial intelligence (AI for science, AI4S). In the meantime, utilizing NLP techniques in geoscience research and practice is wide and convoluted, contributing from knowledge extraction and document classification to question answering and knowledge discovery. In this work, we take the initial step to leverage LLM for science, through a rather straightforward approach. We try to specialize an LLM into geoscience, by further pre-training the model with a vast amount of texts in geoscience, as well as supervised fine-tuning (SFT) the resulting model with our custom collected instruction tuning dataset. These efforts result in a model GeoGalactica consisting of 30 billion parameters. To our best knowledge, it is the largest language model for the geoscience domain. More specifically, GeoGalactica is from further pre-training of Galactica. We train GeoGalactica over a geoscience-related text corpus containing 65 billion tokens, preserving as the largest geoscience-specific text corpus. Then we fine-tune the model with 1 million pairs of instruction-tuning data consisting of questions that demand professional geoscience knowledge to answer. In this technical report, we will illustrate in detail all aspects of GeoGalactica, including data collection, data cleaning, base model selection, pre-training, SFT, and evaluation. We open-source our data curation tools and the checkpoints of GeoGalactica during the first 3/4 of pre-training.2023-12-31T09:22:54ZZhouhan LinCheng DengLe ZhouTianhang ZhangYi XuYutong XuZhongmou HeYuanyuan ShiBeiya DaiYunchong SongBoyi ZengQiyuan ChenYuxun MiaoBo XueShu WangLuoyi FuWeinan ZhangJunxian HeYunqiang ZhuXinbing WangChenghu Zhouhttp://arxiv.org/abs/2401.00445v1Energy-Efficient Power Control for Multiple-Task Split Inference in UAVs: A Tiny Learning-Based Approach2023-12-31T10:16:59ZThe limited energy and computing resources of unmanned aerial vehicles (UAVs) hinder the application of aerial artificial intelligence. The utilization of split inference in UAVs garners significant attention due to its effectiveness in mitigating computing and energy requirements. However, achieving energy-efficient split inference in UAVs remains complex considering of various crucial parameters such as energy level and delay constraints, especially involving multiple tasks. In this paper, we present a two-timescale approach for energy minimization in split inference, where discrete and continuous variables are segregated into two timescales to reduce the size of action space and computational complexity. This segregation enables the utilization of tiny reinforcement learning (TRL) for selecting discrete transmission modes for sequential tasks. Moreover, optimization programming (OP) is embedded between TRL's output and reward function to optimize the continuous transmit power. Specifically, we replace the optimization of transmit power with that of transmission time to decrease the computational complexity of OP since we reveal that energy consumption monotonically decreases with increasing transmission time. The replacement significantly reduces the feasible region and enables a fast solution according to the closed-form expression for optimal transmit power. Simulation results show that the proposed algorithm can achieve a higher probability of successful task completion with lower energy consumption.2023-12-31T10:16:59ZChenxi ZhaoMin ShengJunyu LiuTianshu ChuJiandong Lihttp://arxiv.org/abs/2401.00446v1Dissipation of AGN jets in a clumpy interstellar medium2023-12-31T10:29:38ZAccreting supermassive black holes (SMBHs) frequently power jets that interact with the interstellar/circumgalactic medium (ISM/CGM), regulating star-formation in the galaxy. Highly supersonic jets launched by active galactic nuclei (AGN) power a cocoon that confines them and shocks the ambient medium. We build upon the models of narrow conical jets interacting with a smooth ambient medium, to include the effect of dense clouds that are an essential ingredient of a multiphase ISM. The key physical ingredient of this model is that the clouds along the supersonic jet-beam strongly decelerate the jet-head, but the subsonic cocoon easily moves around the clouds without much resistance. We propose scalings for important physical quantities -- cocoon pressure, head & cocoon speed, and jet radius. We obtain, for the first time, the analytic condition on clumpiness of the ambient medium for the jet to dissipate within the cocoon and verify it with numerical simulations of conical jets interacting with a uniform ISM with embedded spherical clouds. A jet is defined to be dissipated when the cocoon speed exceeds the speed of the jet-head. We compare our models to more sophisticated numerical simulations, direct observations of jet-ISM interaction (e.g., quasar J1316+1753), and discuss implications for the Fermi/eROSITA bubbles. Our work also motivates effective subgrid models for AGN jet feedback in a clumpy ISM unresolved by the present generation of cosmological galaxy formation simulations.2023-12-31T10:29:38Z23 pages, 12 figures, 3 tables; to be submitted; comments are welcome; accompanying video: http://youtu.be/DUpSwMMrGfkRiju DuttaPrateek SharmaKartick C. SarkarJames M. Stonehttp://arxiv.org/abs/2401.00470v2More on $G$-flux and General Hodge Cycles on the Fermat Sextic2024-02-26T13:32:24ZWe study M-Theory solutions with $G$-flux on the Fermat sextic Calabi-Yau fourfold, focussing on the relationship between the number of stabilized complex structure moduli and the tadpole contribution of the flux. We use two alternative approaches to define the fluxes: algebraic cycles and (appropriately quantized) Griffiths residues. In both cases, we collect evidence for the non-existence of solutions which stabilize all moduli and stay within the tadpole bound2023-12-31T11:58:14Zv2: typos corrected and references addedAndreas P. BraunHugo FortinDaniel Lopez GarciaRoberto Villaflor Loyolahttp://arxiv.org/abs/2401.00472v1On Thurston's geometrical space form problem: on quasi space forms2023-12-31T12:03:20ZA proposal is made for what may well be the most elementary Riemannian spaces which are homogeneous but not isotropic. In other words: a proposal is made for what may well be the nicest symmetric spaces beyond the real space forms, that is, beyond the Riemannian spaces which are homogeneous and isotropic. The above qualification of `'nicest symmetric spaces'' finds a justification in that, together with the real space forms, these spaces are most natural with respect to the importance in human vision of our ability to readily recognise conformal things and in that these spaces are most natural with respect to what in Weyl's view is symmetry in Riemannian geometry.
Following his suggestion to remove the real space forms' isotropy condition, the quasi space forms thus introduced do offer a metrical, local geometrical solution to the geometrical space form problem as posed by Thurston in his 1979 Princeton Lecture Notes on `'The Geometry and Topology of 3-manifolds''. Roughly speaking, quasi space forms are the Riemannian manifolds of dimension greater than or equal to 3, which are not real space forms but which admit two orthogonally complementary distributions such that at all points all the 2-planes that in the tangent spaces there are situated in a same position relative to these distributions do have the same sectional curvatures.2023-12-31T12:03:20Z20 pagesStefan HaesenMiroslava Petrović-TorgaševLeopold Verstraelenhttp://arxiv.org/abs/2401.00482v1Structural deformation and irreversible magnetic properties of flexible Co/Pt and Co/Pd thin films2023-12-31T12:55:59ZThe successful commercialization of flexible spintronic devices requires a complete understanding of the impact of external strain on the structural, electronic, and magnetic properties of a system. The impact of bending-induced strain on flexible films is studied quite well. However, little is known about the effect of other modes of flexibility, e.g., wrinkling, twisting, peeling, and stretching on the functional properties of flexible films. In this context, perpendicular magnetic anisotropic Co/Pt and Co/Pd thin films are prepared on flexible Kapton substrates, and the impact of the peeling mode is studied in detail. The peeling method generates numerous cracks, and buckling in the thin film, along with localized blister formation imaged by scanning electron microscopy. Further, the resistivity measurement confirms a significant enhancement in sample resistance owing to the severe damage of the films. The structural discontinuities strongly affect the magnetization reversal phenomena as measured by the magneto-optic Kerr effect (MOKE)-based microscopy. The bubble domains got converted to elongated-shaped domains due to several hindrances to the wall motion after strain application. Further, the relaxation measurements reveal that the thermal energy is insufficient to switch the magnetization at a few areas due to their high pinning potential associated with the damages. In contrast to bending-induced strain, here, all the modifications in the functional properties are found to be irreversible in nature.2023-12-31T12:55:59ZEsita PandeyShaktiranjan MohantyAbhisek MishraBhuvneshwari SharmaSubhankar Bedantahttp://arxiv.org/abs/2401.00484v1Directional flow in perivascular networks: Mixed finite elements for reduced-dimensional models on graphs2023-12-31T12:57:35ZThe flow of cerebrospinal fluid through the perivascular spaces of the brain is believed to play a crucial role in eliminating toxic waste proteins. While the driving forces of this flow have been enigmatic, experiments have shown that arterial wall motion is central. In this work, we present a network model for simulating pulsatile fluid flow in perivascular networks. We establish the well-posedness of this model in the primal and dual mixed variational settings, and show how it can be discretized using mixed finite elements. Further, we utilize this model to investigate fundamental questions concerning the physical mechanisms governing perivascular fluid flow. Notably, our findings reveal that arterial pulsations can induce directional flow in branching perivascular networks.2023-12-31T12:57:35ZIngeborg G. GjerdeMiroslav KuchtaMarie E. RognesBarbara Wohlmuthhttp://arxiv.org/abs/2401.00486v1Molecular Hybridization Induced Antidamping and Sizable Enhanced Spin-to-Charge Conversion in Co20Fe60B20/$β$-W/C60 Heterostructures2023-12-31T12:59:23ZDevelopment of power efficient spintronics devices has been the compelling need in the post-CMOS technology era. The effective tunability of spin-orbit-coupling (SOC) in bulk and at the interfaces of hybrid materials stacking is a prerequisite for scaling down the dimension and power consumption of these devices. In this work, we demonstrate the strong chemisorption of C60 molecules when grown on the high SOC $β$-W layer. The parent CFB/$β$-W bilayer exhibits large spin-to-charge interconversion efficiency, which can be ascribed to the interfacial SOC observed at the Ferromagnet/Heavy metal interface. Further, the adsorption of C60 molecules on $β$-W reduces the effective Gilbert damping by $\sim$15% in the CFB/$β$-W/C60 heterostructures. The anti-damping is accompanied by a gigantic $\sim$115% enhancement in the spin-pumping induced output voltage owing to the molecular hybridization. The non-collinear Density Functional Theory calculations confirm the long-range enhancement of SOC of $β$-W upon the chemisorption of C60 molecules, which in turn can also enhance the SOC at the CFB/$β$-W interface in CFB/$β$-W/C60 heterostructures. The combined amplification of bulk as well interfacial SOC upon molecular hybridization stabilizes the anti-damping and enhanced spin-to-charge conversion, which can pave the way for the fabrication of power efficient spintronics devices.2023-12-31T12:59:23ZAntarjami SahooAritra MukhopadhyayaSwayang Priya MahantaMd. Ehesan AliSubhankar Bedantahttp://arxiv.org/abs/2401.00491v1On the sense of convergence in the dyadic representation theorem2023-12-31T13:21:37ZThe dyadic representation of any singular integral operator, as an average of dyadic model operators, has found many applications. While for many purposes it is enough to have such a representation for a "suitable class" of test functions, we show that, under quite general assumptions (essentially minimal ones to make sense of the formula), the representation is actually valid for all pairs $(f,g)\in L^p(\mathbb R^d)\times L^{p'}(\mathbb R^d)$, not just test functions.2023-12-31T13:21:37Z25 pagesTuomas Hytönenhttp://arxiv.org/abs/2401.00496v2SAR-RARP50: Segmentation of surgical instrumentation and Action Recognition on Robot-Assisted Radical Prostatectomy Challenge2024-01-23T23:30:57ZSurgical tool segmentation and action recognition are fundamental building blocks in many computer-assisted intervention applications, ranging from surgical skills assessment to decision support systems. Nowadays, learning-based action recognition and segmentation approaches outperform classical methods, relying, however, on large, annotated datasets. Furthermore, action recognition and tool segmentation algorithms are often trained and make predictions in isolation from each other, without exploiting potential cross-task relationships. With the EndoVis 2022 SAR-RARP50 challenge, we release the first multimodal, publicly available, in-vivo, dataset for surgical action recognition and semantic instrumentation segmentation, containing 50 suturing video segments of Robotic Assisted Radical Prostatectomy (RARP). The aim of the challenge is twofold. First, to enable researchers to leverage the scale of the provided dataset and develop robust and highly accurate single-task action recognition and tool segmentation approaches in the surgical domain. Second, to further explore the potential of multitask-based learning approaches and determine their comparative advantage against their single-task counterparts. A total of 12 teams participated in the challenge, contributing 7 action recognition methods, 9 instrument segmentation techniques, and 4 multitask approaches that integrated both action recognition and instrument segmentation. The complete SAR-RARP50 dataset is available at: https://rdr.ucl.ac.uk/projects/SARRARP50_Segmentation_of_surgical_instrumentation_and_Action_Recognition_on_Robot-Assisted_Radical_Prostatectomy_Challenge/1910912023-12-31T13:32:18ZDimitrios PsychogyiosEmanuele ColleoniBeatrice Van AmsterdamChih-Yang LiShu-Yu HuangYuchong LiFucang JiaBaosheng ZouGuotai WangYang LiuMaxence BoelsJiayu HuoRachel SparksProkar DasguptaAlejandro GranadosSebastien OurselinMengya XuAn WangYanan WuLong BaiHongliang RenAtsushi YamadaYuriko HaraiYuto IshikawaKazuyuki HayashiJente SimoensPieter DeBackerFrancesco CisterninoGabriele FurnariAlex MottrieFederica FerragutiSatoshi KondoSatoshi KasaiKousuke HirasawaSoohee KimSeung Hyun LeeKyu Eun LeeHyoun-Joong KongKui FuChao LiShan AnStefanie KrellSebastian BodenstedtNicolas AyobiAlejandra PerezSantiago RodriguezJuanita PuentesPablo ArbelaezOmid MohareriDanail Stoyanovhttp://arxiv.org/abs/2401.00401v1Multiplayer Battle Game-Inspired Optimizer for Complex Optimization Problems2023-12-31T05:28:12ZVarious popular multiplayer battle royale games share a lot of common elements. Drawing from our observations, we summarized these shared characteristics and subsequently proposed a novel heuristic algorithm named multiplayer battle game-inspired optimizer (MBGO). The proposed MBGO streamlines mainstream multiplayer battle royale games into two discrete phases: movement and battle. Specifically, the movement phase incorporates the principles of commonly encountered ``safe zones'' to incentivize participants to relocate to areas with a higher survival potential. The battle phase simulates a range of strategies adopted by players in various situations to enhance the diversity of the population. To evaluate and analyze the performance of the proposed MBGO, we executed it alongside eight other algorithms, including three classics and five latest ones, across multiple diverse dimensions within the CEC2017 and CEC2020 benchmark functions. In addition, we employed several industrial design problems to evaluate the scalability and practicality of the proposed MBGO. The results of the statistical analysis reveal that the novel MBGO demonstrates significant competitiveness, excelling not only in convergence speed, but also in achieving high levels of convergence accuracy across both benchmark functions and real-world problems.2023-12-31T05:28:12ZYuefeng XuRui ZhongChao ZhangJun Yuhttp://arxiv.org/abs/2401.00489v2Hodge Theory of Abelian Covers2024-07-17T14:45:23ZMotivated by classical Alexander invariants of affine hypersurface complements, we endow certain finite dimensional quotients of the homology of abelian covers of complex algebraic varieties with a canonical and functorial mixed Hodge structure (MHS). More precisely, we focus on covers which arise algebraically in the following way: if $U$ is a smooth connected complex algebraic variety and $G$ is a complex semiabelian variety, the pullback of the exponential map by an algebraic morphism $f:U\to G$ yields a covering space $π:U^f\to U$ whose group of deck transformations is $π_1(G)$. The new MHSs are compatible with Deligne's MHS on the homology of $U$ through the covering map $π$ and satisfy a direct sum decomposition as MHSs into generalized eigenspaces by the action of deck transformations. This provides a vast generalization of the previous results regarding univariable Alexander modules by Geske, Maxim, Wang and the authors. Lastly, we reduce the problem of whether the first Betti number of the Milnor fiber of a central hyperplane arrangement complement is combinatorial to a question about the Hodge filtration of certain MHSs defined in this paper, providing evidence that the new structures contain interesting information.2023-12-31T13:18:06Z72 pagesEva ElduqueMoisés Herradón Cuetohttp://arxiv.org/abs/2401.00407v2CMB lensing from early-formed dark matter halos2024-05-20T01:11:30ZSome theoretical models for the early Universe predict a spike-type enhancement in the primordial power spectrum on a small scale, which would result in forming early-formed dark matter halos~(EFHs). Some recent studies have claimed to have placed limits on such small scales, which, however, involve uncertainties, such as the physics of substructures and the halo-galaxy relations. In this work, we study the cosmic microwave background~(CMB) lensing effect, considering the existence of EFHs, and investigate the potential to probe the EFHs and the primordial perturbations on scales smaller than $1\mathrm{Mpc}$, complementing these previous studies. We numerically calculate the angular power spectrum of the lensing potential and the lensed CMB anisotropy of temperature, E-mode, and B-mode polarization, including the nonlinear effects of EFHs. We find the possibility that the lensed CMB temperature anisotropy is significantly enhanced on small scales, $\ell>1000$, and could be tested by component decomposition of observed signals through multifrequency observations. Through the calculation with different models of the spiky-type power spectrum, we demonstrate that the accurate measurements of the CMB lensing effect would provide insight into the abundance of EFHs within the limited mass range around $10^{11}~M_\odot$ and the primordial power spectrum on the limited scales around $k\sim 1\mathrm{Mpc}^{-1}$. In particular, we find that the existence of such EFHs can amplify the lensed anisotropy of CMB B-mode polarization even on large scales, $\ell <100$, as the overall enhancement by $\sim 5 \%$ level compared to the standard structure formation model without EFHs. Therefore, future CMB measurements, such as the LiteBIRD satellite, can probe the existence of the EFHs and the spike-type primordial power spectrum through the precise measurement of the large-scale CMB B-mode polarization.2023-12-31T05:47:18Z14 pages, 7 figuresPhys. Rev. D 109, (2024) 103524Katsuya T. AbeHiroyuki Tashiro10.1103/PhysRevD.109.103524http://arxiv.org/abs/2401.00464v3A note on the Lp-Sobolev inequality2024-11-11T13:49:10ZThe usual Sobolev inequality in $\mathbb{R}^N$, asserts that $\|\nabla u\|_{L^p(\mathbb{R}^N)} \geq \mathcal{S}\|u\|_{L^{p^*}(\mathbb{R}^N)}$ for $1<p<N$ and $p^*=\frac{pN}{N-p}$, with $\mathcal{S}$ being the sharp constant. Based on a recent work of Figalli and Zhang [Duke Math. J., 2022], a weak norm remainder term of Sobolev inequality in a subdomain $Ω\subset \mathbb{R}^N$ with finite measure is established, i.e., for $\frac{2N}{N+1}<p<N$ there exists a constant $\mathcal{C}>0$ independent of $Ω$ such that \[ \|\nabla u\|^p_{L^p(Ω)}
-\mathcal{S}^p\|u\|^p_{L^{p^*}(Ω)} \geq \mathcal{C}|Ω|^{-\fracγ{p^*(p-1)}} \|u\|_{L^{\bar{p}}_w(Ω)}^γ\| u\|_{L^{p^*}(Ω)}^{p-γ},\quad \mbox{for all}\ u\in C^\infty_0(Ω)\setminus\{0\}, \] where $γ=\max\{2,p\}$, $\bar{p}=p^*(p-1)/p$, and $\|\cdot\|_{L^{\bar{p}}_w(Ω)}$ denotes the weak $L^{\bar{p}}$-norm. Moreover, we establish a sharp upper bound of Sobolev inequality in $\mathbb{R}^N$.2023-12-31T11:39:03ZShengbing DengXingliang Tianhttp://arxiv.org/abs/2401.00455v2Worldtube puncture scheme for first- and second-order self-force calculations in the Fourier domain2024-10-26T18:56:05ZSecond-order gravitational self-force theory has recently led to the breakthrough calculation of ``first post-adiabatic'' (1PA) compact-binary waveforms [Phys. Rev. Lett. 130, 241402 (2023)]. The computations underlying those waveforms depend on a method of solving the perturbative second-order Einstein equation on a Schwarzschild background in the Fourier domain. In this paper we present that method, which involves dividing the domain into several regions. Different regions utilize different time slicings and allow for the use of ``punctures'' to tame sources and enforce physical boundary conditions. We demonstrate the method for Lorenz-gauge and Teukolsky equations in the relatively simple case of calculating parametric derivatives (``slow time derivatives'') of first-order fields, which are an essential input at second order.2023-12-31T11:15:41Z41 pages, 10 figures. Minor changes in response to refereePhys. Rev. D 109, 104010 (2024)Jeremy MillerBenjamin LeatherAdam PoundNiels Warburtonhttp://arxiv.org/abs/2401.00480v3Modelling contagious viral dynamics: a kinetic approach based on mutual utility2024-02-09T11:04:16ZThe temporal evolution of a contagious viral disease is modelled as the dynamic progression of different classes of population with individuals interacting pairwise. This interaction follows a binary mechanism typical of kinetic theory, wherein agents aim to improve their condition with respect to a mutual utility target. To this end, we introduce kinetic equations of Boltzmann-type to describe the time evolution of the probability distributions of the multi-agent system. The interactions between agents are defined using principles from price theory, specifically employing Cobb-Douglas utility functions for binary exchange and the Edgeworth box to depict the common exchange area where utility increases for both agents. Several numerical experiments presented in the paper highlight the significance of this mechanism in driving the phenomenon toward endemicity.2023-12-31T12:46:08ZMath. Biosci. Eng. 21 (2024) 4241-4268Giulia BertagliaLorenzo PareschiGiuseppe Toscani10.3934/mbe.2024187http://arxiv.org/abs/2401.00498v3Analytical Model for Atomic Relaxation in Twisted Moiré Materials2024-12-31T16:01:31ZBy virtue of being atomically thin, the electronic properties of heterostructures built from two-dimensional materials are strongly influenced by atomic relaxation. The atomic layers behave as flexible membranes rather than rigid crystals. Here we develop an analytical theory of lattice relaxation in twisted moiré materials. We obtain analytical results for the lattice displacements and corresponding pseudo gauge fields, as a function of twist angle. We benchmark our results for twisted bilayer graphene and twisted WSe$_2$ bilayers using large-scale molecular dynamics simulations. Our \textit{single-parameter} theory is valid in graphene bilayers for twist angles $θ~\gtrsim 0.7^\circ$, and in twisted WSe$_2$ for $θ~\gtrsim 1.6^\circ$. We also investigate how relaxation alters the electronic structure in twisted bilayer graphene, providing a simple extension to the continuum model to account for lattice relaxation.2023-12-31T13:37:17ZPhys. Rev. Lett. 133, 266201 (2024)Mohammed M. Al EzziGayani N. PallewelaChristophe De BeuleE. J. MeleShaffique Adam10.1103/PhysRevLett.133.266201http://arxiv.org/abs/2401.00500v4Deformation Quantization with Separation of Variables of $G_{2,4}(\mathbb{C})$2025-07-23T05:16:29ZWe construct a deformation quantization with separation of variables of the Grassmannian $G_{2,4}(\mathbb{C})$. A star product on $G_{2,4}(\mathbb{C})$ can be explicitly determined as the solution of the recurrence relations for $G_{2,4}(\mathbb{C})$ given by Hara and one of the authors (A. Sako). To provide the solution to the recurrence relations, it is necessary to solve a system of linear equations in each order. However, to give a concrete expression of the general term is not simple because the variables increase with the order of the differentiation of the star product. For this reason, there has been no formula to express the general term of the recurrence relations. In this paper, we overcome this problem by transforming the recurrence relations into simpler ones. We solve the recurrence relations using creation and annihilation operators on a Fock space. From this solution, we obtain an explicit formula of a star product with separation of variables on $G_{2,4}(\mathbb{C})$.2023-12-31T13:46:54ZSIGMA 21 (2025), 061, 32 pagesTaika OkudaAkifumi Sako10.3842/SIGMA.2025.061http://arxiv.org/abs/2401.00454v2Quantum and Classical Communication Complexity of Permutation-Invariant Functions2025-10-13T15:29:16ZThis paper gives a nearly tight characterization of the quantum communication complexity of the permutation-invariant Boolean functions. With such a characterization, we show that the quantum and randomized communication complexity of the permutation-invariant Boolean functions are quadratically equivalent (up to a logarithmic factor). Our results extend a recent line of research regarding query complexity \cite{AA14, Cha19, BCG+20} to communication complexity, showing symmetry prevents exponential quantum speedups.
Furthermore, we show the Log-rank Conjecture holds for any non-trivial total permutation-invariant Boolean function. Moreover, we establish a relationship between the quantum/classical communication complexity and the approximate rank of permutation-invariant Boolean functions. This implies the correctness of the Log-approximate-rank Conjecture for permutation-invariant Boolean functions in both randomized and quantum settings (up to a logarithmic factor).2023-12-31T11:07:49Zreference correctionZiyi GuanYunqi HuangPenghui YaoZekun Yehttp://arxiv.org/abs/2401.00436v6Diff-PCR: Diffusion-Based Correspondence Searching in Doubly Stochastic Matrix Space for Point Cloud Registration2026-04-09T06:05:55ZEfficiently identifying accurate correspondences between point clouds is crucial for both rigid and non-rigid point cloud registration. Existing methods usually rely on geometric or semantic feature embeddings to establish correspondences and then estimate transformations or flow fields. Recently, several state-of-the-art methods have adopted RAFT-like iterative updates to refine solutions. However, these methods still have two major limitations. First, their iterative refinement mechanism lacks transparency, and the update trajectory is largely fixed once the refinement starts, which may lead to suboptimal solutions. Second, they overlook the importance of explicitly refining the correspondence matrix before solving for transformations or flow fields. Most existing approaches compute candidate correspondences in feature space and project the resulting matching matrix only once by using Sinkhorn or dual-softmax normalization. Such a one-shot projection can be far from the globally optimal solution, and these methods usually do not model the distribution of the target matching matrix. In this paper, we propose a novel framework that exploits a denoising diffusion model to predict a search gradient for the optimal matching matrix in doubly stochastic matrix space. Specifically, the diffusion model learns a denoising direction, and the reverse denoising process iteratively searches for improved solutions along this learned direction, which approximates the maximum-likelihood direction of the target matching matrix. To improve efficiency, we design a lightweight denoising module and adopt the accelerated sampling strategy of the Denoising Diffusion Implicit Model (DDIM)\cite{song2020denoising}. Experimental results on 3DMatch/3DLoMatch and 4DMatch/4DLoMatch demonstrate the effectiveness of the proposed framework.2023-12-31T09:24:28ZHaihua ShiQianliang Wuhttp://arxiv.org/abs/2401.00406v1Low-cost Geometry-based Eye Gaze Detection using Facial Landmarks Generated through Deep Learning2023-12-31T05:45:22ZIntroduction: In the realm of human-computer interaction and behavioral research, accurate real-time gaze estimation is critical. Traditional methods often rely on expensive equipment or large datasets, which are impractical in many scenarios. This paper introduces a novel, geometry-based approach to address these challenges, utilizing consumer-grade hardware for broader applicability. Methods: We leverage novel face landmark detection neural networks capable of fast inference on consumer-grade chips to generate accurate and stable 3D landmarks of the face and iris. From these, we derive a small set of geometry-based descriptors, forming an 8-dimensional manifold representing the eye and head movements. These descriptors are then used to formulate linear equations for predicting eye-gaze direction. Results: Our approach demonstrates the ability to predict gaze with an angular error of less than 1.9 degrees, rivaling state-of-the-art systems while operating in real-time and requiring negligible computational resources. Conclusion: The developed method marks a significant step forward in gaze estimation technology, offering a highly accurate, efficient, and accessible alternative to traditional systems. It opens up new possibilities for real-time applications in diverse fields, from gaming to psychological research.2023-12-31T05:45:22ZEsther Enhui YeJohn Enzhou YeJoseph YeJacob YeRunzhou Yehttp://arxiv.org/abs/2401.00409v1A Two-stream Hybrid CNN-Transformer Network for Skeleton-based Human Interaction Recognition2023-12-31T06:46:46ZHuman Interaction Recognition is the process of identifying interactive actions between multiple participants in a specific situation. The aim is to recognise the action interactions between multiple entities and their meaning. Many single Convolutional Neural Network has issues, such as the inability to capture global instance interaction features or difficulty in training, leading to ambiguity in action semantics. In addition, the computational complexity of the Transformer cannot be ignored, and its ability to capture local information and motion features in the image is poor. In this work, we propose a Two-stream Hybrid CNN-Transformer Network (THCT-Net), which exploits the local specificity of CNN and models global dependencies through the Transformer. CNN and Transformer simultaneously model the entity, time and space relationships between interactive entities respectively. Specifically, Transformer-based stream integrates 3D convolutions with multi-head self-attention to learn inter-token correlations; We propose a new multi-branch CNN framework for CNN-based streams that automatically learns joint spatio-temporal features from skeleton sequences. The convolutional layer independently learns the local features of each joint neighborhood and aggregates the features of all joints. And the raw skeleton coordinates as well as their temporal difference are integrated with a dual-branch paradigm to fuse the motion features of the skeleton. Besides, a residual structure is added to speed up training convergence. Finally, the recognition results of the two branches are fused using parallel splicing. Experimental results on diverse and challenging datasets, demonstrate that the proposed method can better comprehend and infer the meaning and context of various actions, outperforming state-of-the-art methods.2023-12-31T06:46:46ZRuoqi YinJianqin Yinhttp://arxiv.org/abs/2401.00414v1Is It Possible to Backdoor Face Forgery Detection with Natural Triggers?2023-12-31T07:16:10ZDeep neural networks have significantly improved the performance of face forgery detection models in discriminating Artificial Intelligent Generated Content (AIGC). However, their security is significantly threatened by the injection of triggers during model training (i.e., backdoor attacks). Although existing backdoor defenses and manual data selection can mitigate those using human-eye-sensitive triggers, such as patches or adversarial noises, the more challenging natural backdoor triggers remain insufficiently researched. To further investigate natural triggers, we propose a novel analysis-by-synthesis backdoor attack against face forgery detection models, which embeds natural triggers in the latent space. We thoroughly study such backdoor vulnerability from two perspectives: (1) Model Discrimination (Optimization-Based Trigger): we adopt a substitute detection model and find the trigger by minimizing the cross-entropy loss; (2) Data Distribution (Custom Trigger): we manipulate the uncommon facial attributes in the long-tailed distribution to generate poisoned samples without the supervision from detection models. Furthermore, to completely evaluate the detection models towards the latest AIGC, we utilize both state-of-the-art StyleGAN and Stable Diffusion for trigger generation. Finally, these backdoor triggers introduce specific semantic features to the generated poisoned samples (e.g., skin textures and smile), which are more natural and robust. Extensive experiments show that our method is superior from three levels: (1) Attack Success Rate: ours achieves a high attack success rate (over 99%) and incurs a small model accuracy drop (below 0.2%) with a low poisoning rate (less than 3%); (2) Backdoor Defense: ours shows better robust performance when faced with existing backdoor defense methods; (3) Human Inspection: ours is less human-eye-sensitive from a comprehensive user study.2023-12-31T07:16:10ZXiaoxuan HanSonglin YangWei WangZiwen HeJing Donghttp://arxiv.org/abs/2401.00444v1RIS-Enabled Integrated Sensing and Communication for 6G Systems2023-12-31T10:10:22ZThe following paper proposes a new target localization system design using an architecture based on reconfigurable intelligent surfaces (RISs) and passive radars (PRs) for integrated sensing and communications systems. The preamble of the communication signal is exploited in order to perform target sensing tasks, which involve detection and localization. The RIS in this case can aid the PR in sensing targets that are otherwise not seen by the PR itself, due to the many obstacles encountered within the propagation channel. Therefore, this work proposes a localization algorithm tailored for the integrated sensing and communications RIS-aided architecture, which is capable of uniquely positioning targets within the scene. The algorithm is capable of detecting the number of targets along with estimating the position of targets via angles and times of arrival. Our simulation results demonstrate the performance of the localization method in terms of different localization and detection metrics and for increasing RIS sizes.2023-12-31T10:10:22ZIEEE Wireless Communications and Networking Conference, 2024Dexin WangAhmad BazziMarwa Chafiihttp://arxiv.org/abs/2401.00447v1User Clustering for STAR-RIS Assisted Full-Duplex NOMA Communication Systems2023-12-31T10:33:34ZIn contrast to conventional reconfigurable intelligent surface (RIS), simultaneous transmitting and reflecting reconfigurable intelligent surface (STAR-RIS) has been proposed recently to enlarge the serving area from 180o to 360o coverage. This work considers the performance of a STAR-RIS aided full-duplex (FD) non-orthogonal multiple access (NOMA) communication systems. The STAR-RIS is implemented at the cell-edge to assist the cell-edge users, while the cell-center users can communicate directly with a FD base station (BS). We first introduce new user clustering schemes for the downlink and uplink transmissions. Then, based on the proposed transmission schemes closed-form expressions of the ergodic rates in the downlink and uplink modes are derived taking into account the system impairments caused by the self interference at the FD-BS and the imperfect successive interference cancellation (SIC). Moreover, an optimization problem to maximize the total sum-rate is formulated and solved by optimizing the amplitudes and the phase-shifts of the STAR-RIS elements and allocating the transmit power efficiently. The performance of the proposed user clustering schemes and the optimal STAR-RIS design are investigated through numerical results2023-12-31T10:33:34ZarXiv admin note: text overlap with arXiv:2309.15037Abdelhamid SalemKai-Kit WongChan-Byoung ChaeYangyang Zhanghttp://arxiv.org/abs/2401.00463v2Analyzing Local Representations of Self-supervised Vision Transformers2024-03-21T14:57:25ZIn this paper, we present a comparative analysis of various self-supervised Vision Transformers (ViTs), focusing on their local representative power. Inspired by large language models, we examine the abilities of ViTs to perform various computer vision tasks with little to no fine-tuning. We design evaluation framework to analyze the quality of local, i.e.\ patch-level, representations in the context of few-shot semantic segmentation, instance identification, object retrieval and tracking. We discover that contrastive learning based methods like DINO produce more universal patch representations that can be immediately applied for downstream tasks with no parameter tuning, compared to masked image modeling. The embeddings learned using the latter approach, e.g. in masked autoencoders, have high variance features that harm distance-based algorithms, such as k-NN, and do not contain useful information for most downstream tasks. Furthermore, we demonstrate that removing these high-variance features enhances k-NN for MAE, as well as for its recent extension Scale-MAE. Finally, we find an object instance retrieval setting where DINOv2, a model pretrained on two orders of magnitude more data, falls short of its less compute intensive counterpart DINO.2023-12-31T11:38:50ZAni VanyanAlvard BarseghyanHakob TamazyanVahan HuroyanHrant KhachatrianMartin Danelljanhttp://arxiv.org/abs/2401.00467v1Amplification of femtosecond pulses with AI-assisted spectral phase modulation2023-12-31T11:45:14ZWe report our investigation on ultrashort laser pulse optimization using an AI algorithm in a system consisting of a mode-locked oscillator, a spectral phase shaper, and a highly nonlinear amplifier. We analyzed the performance of the pulse optimization process as a function of two main parameters: the resolution of spectral phase modulation and the number of agents in the algorithm. We showed that the algorithm could find an optimum phase profile for the seed pulse, which allowed for a reduction of the FWHM of the amplified pulse by 10 fs (from 46 to 36 fs), and significantly reduced the intensity of the side-pulse by a factor of 4.6. Importantly, the algorithm used does not require any training and optimizes the pulse shape without any knowledge about the input pulse parameters or the parameters of the amplifier. We believe the proposed system might be a convenient test bed for evaluating various AI-based algorithms in a pulse optimization task.2023-12-31T11:45:14Z10 pages, 8 figuresMikołaj KrakowskiAlicja KwaśnyGrzegorz Sobońhttp://arxiv.org/abs/2401.00468v1Blockchain and Deep Learning-Based IDS for Securing SDN-Enabled Industrial IoT Environments2023-12-31T11:49:42ZThe industrial Internet of Things (IIoT) involves the integration of Internet of Things (IoT) technologies into industrial settings. However, given the high sensitivity of the industry to the security of industrial control system networks and IIoT, the use of software-defined networking (SDN) technology can provide improved security and automation of communication processes. Despite this, the architecture of SDN can give rise to various security threats. Therefore, it is of paramount importance to consider the impact of these threats on SDN-based IIoT environments. Unlike previous research, which focused on security in IIoT and SDN architectures separately, we propose an integrated method including two components that work together seamlessly for better detecting and preventing security threats associated with SDN-based IIoT architectures. The two components consist in a convolutional neural network-based Intrusion Detection System (IDS) implemented as an SDN application and a Blockchain-based system (BS) to empower application layer and network layer security, respectively. A significant advantage of the proposed method lies in jointly minimizing the impact of attacks such as command injection and rule injection on SDN-based IIoT architecture layers. The proposed IDS exhibits superior classification accuracy in both binary and multiclass categories.2023-12-31T11:49:42ZSamira Kamali PoorazadChafika BenzaıdTarik Talebhttp://arxiv.org/abs/2401.00476v1Reproductive outcome in female wistar rats treated with nhexane, dichloromethane and aqueous ethanol extracts of Cucurbita pepo seed2023-12-31T12:32:05ZIn developing countries, healthcare challenges and expensive infertility treatments has resulted in resurgent interest in medicinal plants. This study was designed to determine if Curcubita pepo seed can enhance female fertility, by assessing the reproductive outcome in female wistar rats treated with n-hexane (nHE), dichloromethane (DCM) and aqueous ethanol (Aq. Eth) extracts of Curcubita pepo seed. Total of 48 rats randomly grouped into 12 (n=4), were treated for 21 days by oral gavage as follows: A (control) = 0.5ml 20% tween 80 (vehicle); B (positive control) = 10mg/kg clomiphene citrate, C, D & E = 142.86, 285.71 and 428.57 mg/kg nHE; F, G & H = 142.86, 285.71 and 428.57 mg/kg DCM ; and I, J & K =142.86, 285.71 and 428.57 mg/kg Aq.Eth extracts. Group L (positive control 2) = 10mg/kg clomiphene citrate for 8 days. Following treatment, the rats were paired with males for mating, designating the confirmation day as gestational day 0 (GD 0). On GD 20, the animals were laparatomised and reproductive outcome was determined by assessing foetal weight, foetal crown-rump length, litter size, number of implantation and resorption sites. Results showed all extracts had no significant (p >0.05) effect on the reproductive outcome indices. Clomiphene citrate significantly decreased reproductive outcome indices. In conclusion, Cucurbita pepo seed did not enhance the reproductive outcome of treated female rats at the doses and duration used in this study. This finding may serve as a springboard for future studies exploring the effect of C.pepo at different doses or durations.2023-12-31T12:32:05Z5 tables, 2 figures AnyanwuC. F GeorgewillO. A. ObinnaVictoria Chttp://arxiv.org/abs/2401.00490v2Kernel Density Estimation for Multiclass Quantification2024-01-02T19:52:24ZSeveral disciplines, like the social sciences, epidemiology, sentiment analysis, or market research, are interested in knowing the distribution of the classes in a population rather than the individual labels of the members thereof. Quantification is the supervised machine learning task concerned with obtaining accurate predictors of class prevalence, and to do so particularly in the presence of label shift. The distribution-matching (DM) approaches represent one of the most important families among the quantification methods that have been proposed in the literature so far. Current DM approaches model the involved populations by means of histograms of posterior probabilities. In this paper, we argue that their application to the multiclass setting is suboptimal since the histograms become class-specific, thus missing the opportunity to model inter-class information that may exist in the data. We propose a new representation mechanism based on multivariate densities that we model via kernel density estimation (KDE). The experiments we have carried out show our method, dubbed KDEy, yields superior quantification performance with respect to previous DM approaches. We also investigate the KDE-based representation within the maximum likelihood framework and show KDEy often shows superior performance with respect to the expectation-maximization method for quantification, arguably the strongest contender in the quantification arena to date.2023-12-31T13:19:27Zfixed broken references to appendicesAlejandro MoreoPablo GonzálezJuan José del Cozhttp://arxiv.org/abs/2401.00494v1Generalization of the Bargmann-Wigner approach to constructing relativistic fields2023-12-31T13:28:28ZWe review the method for constructing local relativistic fields corresponding to the Bargmann-Wigner wave functions that describe the unitary irreducible representations of the $4D$ Poincaré group. The method is based on the use of the generalized Wigner operator connecting the wave functions of induced representations and local relativistic fields. Applications of this operator for constructing massive local relativistic fields as well as massless helicity local fields and massless local infinite spin fields are considered.2023-12-31T13:28:28Z1+12 pages, Contribution to the Proceedings of the International Conference on Particle Physics and Cosmology (professor V.A. Rubakov memorial conference), October 02-07, 2023, Yerevan, ArmeniaI. L. BuchbinderS. A. FedorukA. P. IsaevM. A. Podoinitsynhttp://arxiv.org/abs/2401.00495v1Lattice construction of mixed 't Hooft anomaly with higher-form symmetry2023-12-31T13:30:48ZIn this talk, we give the lattice regularized formulation of the mixed 't Hooft anomaly between the $\mathbb{Z}_N$ $1$-form symmetry and the $θ$ periodicity for $4$d pure Yang-Mills theory, which was originally discussed by Gaiotto $\textit{et al.}$ in the continuum description. For this purpose, we define the topological charge of the lattice $SU(N)$ gauge theory coupled with the background $\mathbb{Z}_N$ $2$-form gauge fields $B_p$ by generalizing Lüscher's construction of the $SU(N)$ topological charge. We show that this lattice topological charge enjoys the fractional $1/N$ shift completely characterized by the background gauge field $B_p$, and this rigorously proves the mixed 't Hooft anomaly with the finite lattice spacings. As a consequence, the Yang-Mills vacua at $θ$ and $θ+2π$ are distinct as the symmetry-protected topological states when the confinement is assumed.2023-12-31T13:30:48Z8 pages, 2 figures, talk presented at the 40th International Symposium on Lattice Field Theory (Lattice2023), July 31st - August 4th, 2023, Fermi National Accelerator LaboratoryMotokazu AbeOkuto MorikawaSoma OnodaHiroshi SuzukiYuya Tanizakihttp://arxiv.org/abs/2401.00428v3Training toward significance with the decorrelated event classifier transformer neural network2024-07-11T01:50:17ZExperimental particle physics uses machine learning for many tasks, where one application is to classify signal and background events. This classification can be used to bin an analysis region to enhance the expected significance for a mass resonance search. In natural language processing, one of the leading neural network architectures is the transformer. In this work, an event classifier transformer is proposed to bin an analysis region, in which the network is trained with special techniques. The techniques developed here can enhance the significance and reduce the correlation between the network's output and the reconstructed mass. It is found that this trained network can perform better than boosted decision trees and feed-forward networks.2023-12-31T08:57:29Z11 pages, 7 figures, 1 tablePhys. Rev. D 109, 096035 (2024)Jaebak Kimhttp://arxiv.org/abs/2401.00412v2Toward the theoretically observable limit of electron density distribution by single-crystal synchrotron X-ray diffraction: The case of orbitally ordered Ti-3d^1 in YTiO_32024-07-08T16:43:10ZThe theoretically observable limit of electron density distribution by single-crystal X-ray diffraction is discussed. When F_{orb} and δF are defined as, respectively, the partial structure factor for an orbital and the deviation of the observed F from the true F, the accuracy of electron density attributable to F_{orb} is chiefly determined by the number of reflections satisfying the condition F_{orb}/F > δF/F. Since F_{orb}/F, which is generally small for crystals with large F(0,0,0), is constant under a given set of experimental conditions, δF/F must be reduced to increase the number of reflections satisfying F_{orb}/F > δF/F. The present paper demonstrates how to reduce δF mathematically and experimentally, and the following topics are covered: the Poisson statistics, accumulation of errors in the data collection and reduction procedure, multiple diffraction, conversion error from F^2 to F in refinement programs, which is unavoidable when the input quantities have different dimension from F, weighting of reflections, and tips. For demonstration, observation of the electron density of the Ti-3d^1 orbital in YTiO_3 by synchrotron single-crystal X-ray diffraction is presented.2023-12-31T07:08:06Z68 pages, 20 figuresTerutoshi SakakuraYoshihisa IshikawaShunji KishimotoYasuyuki TakenakaKiyoaki TanakaShigeki MiyasakaYoshinori TokuraYukio NodaNobuo IshizawaHajime SagayamaHajime YamamotoHiroyuki Kimurahttp://arxiv.org/abs/2401.00477v2Coding for Gaussian Two-Way Channels: Linear and Learning-Based Approaches2025-04-23T13:16:13ZAlthough user cooperation cannot improve the capacity of Gaussian two-way channels (GTWCs) with independent noises, it can improve communication reliability. In this work, we aim to enhance and balance the communication reliability in GTWCs by minimizing the sum of error probabilities via joint design of encoders and decoders at the users. We first formulate general encoding/decoding functions, where the user cooperation is captured by the coupling of user encoding processes. The coupling effect renders the encoder/decoder design non-trivial, requiring effective decoding to capture this effect, as well as efficient power management at the encoders within power constraints. To address these challenges, we propose two different two-way coding strategies: linear coding and learning-based coding. For linear coding, we propose optimal linear decoding and discuss new insights on encoding regarding user cooperation to balance reliability. We then propose an efficient algorithm for joint encoder/decoder design. For learning-based coding, we introduce a novel recurrent neural network (RNN)-based coding architecture, where we propose interactive RNNs and a power control layer for encoding, and we incorporate bi-directional RNNs with an attention mechanism for decoding. Through simulations, we show that our two-way coding methodologies outperform conventional channel coding schemes (that do not utilize user cooperation) significantly in sum-error performance. We also demonstrate that our linear coding excels at high signal-to-noise ratios (SNRs), while our RNN-based coding performs best at low SNRs. We further investigate our two-way coding strategies in terms of power distribution, two-way coding benefit, different coding rates, and block-length gain.2023-12-31T12:40:18ZThis work has been accepted for publication in the IEEE Transactions on Information TheoryJunghoon KimTaejoon KimAnindya Bijoy DasSeyyedali HosseinalipourDavid J. LoveChristopher G. Brintonhttp://arxiv.org/abs/2401.00473v1Emulating insect brains for neuromorphic navigation2023-12-31T12:05:42ZBees display the remarkable ability to return home in a straight line after meandering excursions to their environment. Neurobiological imaging studies have revealed that this capability emerges from a path integration mechanism implemented within the insect's brain. In the present work, we emulate this neural network on the neuromorphic mixed-signal processor BrainScaleS-2 to guide bees, virtually embodied on a digital co-processor, back to their home location after randomly exploring their environment. To realize the underlying neural integrators, we introduce single-neuron spike-based short-term memory cells with axo-axonic synapses. All entities, including environment, sensory organs, brain, actuators, and the virtual body, run autonomously on a single BrainScaleS-2 microchip. The functioning network is fine-tuned for better precision and reliability through an evolution strategy. As BrainScaleS-2 emulates neural processes 1000 times faster than biology, 4800 consecutive bee journeys distributed over 320 generations occur within only half an hour on a single neuromorphic core.2023-12-31T12:05:42ZKorbinian SchreiberTimo WunderlichPhilipp SpilgerSebastian BillaudelleBenjamin CramerYannik StradmannChristian PehleEric MüllerMihai A. PetroviciJohannes SchemmelKarlheinz Meierhttp://arxiv.org/abs/2401.00457v4Two types of filtrations for $\mathrm{wK4}$ and its relatives2025-06-12T18:14:15ZWe study the finite model property of subframe logics with expressible transitive reflexive closure modality. For $m>0$, let $\mathrm{L}_m$ be the logic defined by axiom $\lozenge^{m+1} p\to \lozenge p\vee p$. We construct filtrations for the logics $\mathrm{L}_m$. It follows that these logics and their tense counterparts have the finite model property. Then we show that every canonical subframe logic that contains $\mathrm{L}_m$ have the finite model property.2023-12-31T11:18:59ZAndrey KudinovIlya Shapirovskyhttp://arxiv.org/abs/2401.00413v2Real-Time FJ/MAC PDE Solvers via Tensorized, Back-Propagation-Free Optical PINN Training2024-01-04T06:25:16ZSolving partial differential equations (PDEs) numerically often requires huge computing time, energy cost, and hardware resources in practical applications. This has limited their applications in many scenarios (e.g., autonomous systems, supersonic flows) that have a limited energy budget and require near real-time response. Leveraging optical computing, this paper develops an on-chip training framework for physics-informed neural networks (PINNs), aiming to solve high-dimensional PDEs with fJ/MAC photonic power consumption and ultra-low latency. Despite the ultra-high speed of optical neural networks, training a PINN on an optical chip is hard due to (1) the large size of photonic devices, and (2) the lack of scalable optical memory devices to store the intermediate results of back-propagation (BP). To enable realistic optical PINN training, this paper presents a scalable method to avoid the BP process. We also employ a tensor-compressed approach to improve the convergence and scalability of our optical PINN training. This training framework is designed with tensorized optical neural networks (TONN) for scalable inference acceleration and MZI phase-domain tuning for \textit{in-situ} optimization. Our simulation results of a 20-dim HJB PDE show that our photonic accelerator can reduce the number of MZIs by a factor of $1.17\times 10^3$, with only $1.36$ J and $1.15$ s to solve this equation. This is the first real-size optical PINN training framework that can be applied to solve high-dimensional PDEs.2023-12-31T07:10:15ZML with New Compute Paradigms (MLNCP) at NeurIPS 2023Yequan ZhaoXian XiaoXinling YuZiyue LiuZhixiong ChenGeza KurczveilRaymond G. BeausoleilZheng Zhanghttp://arxiv.org/abs/2401.00417v2Stability for the 2-D plane Poiseuille flow in finite channel2024-03-03T04:51:57ZIn this paper, we study the stability for 2-D plane Poiseuille flow $(1-y^2,0)$ in a channel $\mathbb{T}\times (-1,1)$ with Navier-slip boundary condition. We prove that if the initial perturbation for velocity field $u_0$ satisfies that $\|u_0\|_{H^{\frac{7}{2}+}} \leq ε_1 ν^{2/3}$ for some suitable small $0<ε_1 \ll 1$ independent of viscosity coefficient $ν$, then the solution to the Navier-Stokes equations is global in time and does not transit from the plane Poiseuille flow. This result improves the result of \cite{DL1} from $3/4$ to $2/3$.2023-12-31T07:44:58ZThis version fixes an error in the proof of precious version, and improves the result of [18] form 3/4 to 2/3 for slip boundary value problem. The case of non-slip boundary value problem is not included in this versionShijin DingZhilin Linhttp://arxiv.org/abs/2401.00423v1MSGNet: Learning Multi-Scale Inter-Series Correlations for Multivariate Time Series Forecasting2023-12-31T08:23:24ZMultivariate time series forecasting poses an ongoing challenge across various disciplines. Time series data often exhibit diverse intra-series and inter-series correlations, contributing to intricate and interwoven dependencies that have been the focus of numerous studies. Nevertheless, a significant research gap remains in comprehending the varying inter-series correlations across different time scales among multiple time series, an area that has received limited attention in the literature. To bridge this gap, this paper introduces MSGNet, an advanced deep learning model designed to capture the varying inter-series correlations across multiple time scales using frequency domain analysis and adaptive graph convolution. By leveraging frequency domain analysis, MSGNet effectively extracts salient periodic patterns and decomposes the time series into distinct time scales. The model incorporates a self-attention mechanism to capture intra-series dependencies, while introducing an adaptive mixhop graph convolution layer to autonomously learn diverse inter-series correlations within each time scale. Extensive experiments are conducted on several real-world datasets to showcase the effectiveness of MSGNet. Furthermore, MSGNet possesses the ability to automatically learn explainable multi-scale inter-series correlations, exhibiting strong generalization capabilities even when applied to out-of-distribution samples.2023-12-31T08:23:24Z13 pages, 12 figuresWanlin CaiYuxuan LiangXianggen LiuJianshuai FengYuankai Wuhttp://arxiv.org/abs/2401.00427v2The functional volume product under heat flow2024-03-20T14:44:53ZWe prove that the functional volume product for even functions is monotone increasing along the Fokker--Planck heat flow. This in particular yields a new proof of the functional Blaschke--Santaló inequality by K. Ball and also Artstein-Avidan--Klartag--Milman in the even case.
This result is the consequence of a new understanding of the regularizing property of the Ornstein--Uhlenbeck semigroup. That is, we establish an improvement of Borell's reverse hypercontractivity inequality for even functions and identify the sharp range of the admissible exponents. As another consequence of successfully identifying the sharp range for the inequality, we derive the sharp $L^p$-$L^q$ inequality for the Laplace transform for even functions. The best constant of the inequality is attained by centered Gaussians, and thus this provides an analogous result to Beckner's sharp Hausdorff--Young inequality.
Our technical novelty in the proof is the use of the Brascamp--Lieb inequality for log-concave measures and Cramér--Rao's inequality in this context.2023-12-31T08:48:33ZIn this update, we have mentioned about the "detropicalised" approach that has been proposed in the discussion of Klartag and Tao in Tao's blog post as it is closely related this work. We have mentioned works of Berndtsson--Mastrantonis--Rubinstein and Kolesnikov--Werner. Also, we have split the result on the stability from this version. This part will be in the forthcoming paperShohei NakamuraHiroshi Tsujihttp://arxiv.org/abs/2401.00452v3Multi-scale cross-attention transformer encoder for event classification2024-02-15T02:43:00ZWe deploy an advanced Machine Learning (ML) environment, leveraging a multi-scale cross-attention encoder for event classification, towards the identification of the $gg\to H\to hh\to b\bar b b\bar b$ process at the High Luminosity Large Hadron Collider (HL-LHC), where $h$ is the discovered Standard Model (SM)-like Higgs boson and $H$ a heavier version of it (with $m_H>2m_h$). In the ensuing boosted Higgs regime, the final state consists of two fat jets. Our multi-modal network can extract information from the jet substructure and the kinematics of the final state particles through self-attention transformer layers. The diverse learned information is subsequently integrated to improve classification performance using an additional transformer encoder with cross-attention heads. We ultimately prove that our approach surpasses in performance current alternative methods used to establish sensitivity to this process, whether solely based on kinematic analysis or else on a combination of this with mainstream ML approaches. Then, we employ various interpretive methods to evaluate the network results, including attention map analysis and visual representation of Gradient-weighted Class Activation Mapping (Grad-CAM). Finally, we note that the proposed network is generic and can be applied to analyse any process carrying information at different scales. Our code is publicly available for generic use.2023-12-31T11:03:28ZTypos correctedA. HammadS. MorettiM. Nojirihttp://arxiv.org/abs/2401.00458v1Two sequences of spiral galaxies with different shapes of the metallicity gradients2023-12-31T11:20:27ZWe considered two sequences of spiral galaxies with different shapes of the radial gas-phase oxygen abundance distributions from the galaxies in the MaNGA survey: (1) Galaxies in which the gradient is well approximated by a single linear relation across the whole disc, that is, galaxies with an S (slope) gradients, (2) galaxies in which the metallicity in the inner region of the disc is at a nearly constant level and the gradient is negative at larger radii, that is, galaxies with level-slope (LS) gradients. We also selected galaxies with a nearly uniform oxygen abundance across the whole galaxy, that is, galaxies with level (L) gradients that can be the final evolutionary stage of the two galaxy sequences described above. The radial nitrogen abundance distributions in galaxies with LS oxygen abundance distributions also show breaks at radii smaller than the O/H distribution breaks. The observed behaviour of the oxygen and nitrogen abundances with radius in these galaxies can be explained by the time delay between the nitrogen and oxygen enrichment together with the variation in the star formation history along the radius. These galaxies clearly show the effect of the inside-out disc evolution model. We find that the shape of the radial abundance distribution in a galaxy is not related to its macroscopic characteristics (rotation velocity, stellar mass, isophotal radius, and star formation rate). The correlations between the gradient slopes and macroscopic characteristics of galaxies are weak in the sense that the scatter of the points in each diagram is large. We also examined the properties of the Milky Way in the context of the considered galaxy samples.2023-12-31T11:20:27Z20 pages, 19 figures, accepted to Astronomy and AstrophysicsL. S. PilyuginG. Tautvaisienehttp://arxiv.org/abs/2401.00485v1Additive spectrum preserving mappings from von Neumann algebras2023-12-31T12:57:45ZWe establish Jafarian's 2009 conjecture that every additive spectrum preserving mapping from a von Neumann algebra onto a semisimple Banach algebra is a Jordan isomorphism.2023-12-31T12:57:45Z13 pagesMartin MathieuFrancois Schulzhttp://arxiv.org/abs/2401.00499v3Generating High-Precision Force Fields for Molecular Dynamics Simulations to Study Chemical Reaction Mechanisms using Molecular Configuration Transformer2024-04-11T17:15:43ZTheoretical studies on chemical reaction mechanisms have been crucial in organic chemistry. Traditionally, calculating the manually constructed molecular conformations of transition states for chemical reactions using quantum chemical calculations is the most commonly used method. However, this way is heavily dependent on individual experience and chemical intuition. In our previous study, we proposed a research paradigm that uses enhanced sampling in molecular dynamics simulations to study chemical reactions. This approach can directly simulate the entire process of a chemical reaction. However, the computational speed limits the use of high-precision potential energy functions for simulations. To address this issue, we present a scheme for training high-precision force fields for molecular modeling using a previously developed graph-neural-network-based molecular model, molecular configuration transformer. This potential energy function allows for highly accurate simulations at a low computational cost, leading to more precise calculations of the mechanism of chemical reactions. We applied this approach to study a Claisen rearrangement reaction and a Carbonyl insertion reaction catalyzed by Manganese.2023-12-31T13:43:41ZSihao YuanXu HanJun ZhangZhaoxin XieCheng FanYunlong XiaoYi Qin GaoYi Isaac Yanghttp://arxiv.org/abs/2401.00435v1Bidirectional Trained Tree-Structured Decoder for Handwritten Mathematical Expression Recognition2023-12-31T09:24:21ZThe Handwritten Mathematical Expression Recognition (HMER) task is a critical branch in the field of OCR. Recent studies have demonstrated that incorporating bidirectional context information significantly improves the performance of HMER models. However, existing methods fail to effectively utilize bidirectional context information during the inference stage. Furthermore, current bidirectional training methods are primarily designed for string decoders and cannot adequately generalize to tree decoders, which offer superior generalization capabilities and structural analysis capacity. In order to overcome these limitations, we propose the Mirror-Flipped Symbol Layout Tree (MF-SLT) and Bidirectional Asynchronous Training (BAT) structure. Our method extends the bidirectional training strategy to the tree decoder, allowing for more effective training by leveraging bidirectional information. Additionally, we analyze the impact of the visual and linguistic perception of the HMER model separately and introduce the Shared Language Modeling (SLM) mechanism. Through the SLM, we enhance the model's robustness and generalization when dealing with visual ambiguity, particularly in scenarios with abundant training data. Our approach has been validated through extensive experiments, demonstrating its ability to achieve new state-of-the-art results on the CROHME 2014, 2016, and 2019 datasets, as well as the HME100K dataset. The code used in our experiments will be publicly available.2023-12-31T09:24:21ZHanbo ChengChenyu LiuPengfei HuZhenrong ZhangJiefeng MaJun Duhttp://arxiv.org/abs/2401.00474v1Probabilistically Checkable Reconfiguration Proofs and Inapproximability of Reconfiguration Problems2023-12-31T12:22:11ZMotivated by the inapproximability of reconfiguration problems, we present a new PCP-type characterization of PSPACE, which we call a probabilistically checkable reconfiguration proof (PCRP): Any PSPACE computation can be encoded into an exponentially long sequence of polynomially long proofs such that every adjacent pair of the proofs differs in at most one bit, and every proof can be probabilistically checked by reading a constant number of bits.
Using the new characterization, we prove PSPACE-completeness of approximate versions of many reconfiguration problems, such as the Maxmin $3$-SAT Reconfiguration problem. This resolves the open problem posed by Ito, Demaine, Harvey, Papadimitriou, Sideri, Uehara, and Uno (ISAAC 2008; Theor. Comput. Sci. 2011) as well as the Reconfiguration Inapproximability Hypothesis by Ohsaka (STACS 2023) affirmatively. We also present PSPACE-completeness of approximating the Maxmin Clique Reconfiguration problem to within a factor of $n^ε$ for some constant $ε> 0$.2023-12-31T12:22:11Z31 pagesProceedings of the 56th Annual ACM Symposium on Theory of Computing (STOC), pp. 1435--1445, 2024Shuichi HiraharaNaoto Ohsaka10.1145/3618260.3649667http://arxiv.org/abs/2401.00462v2On the existence of analytic families of G-stable lattices and their reductions2024-11-18T23:15:33ZIn this article, we prove the existence of rigid analytic families of $G$-stable lattices with locally constant reductions inside families of representations of a topologically compact group $G$, extending a result of Hellman obtained in the semi-simple residual case. Implementing this generalization in the context of Galois representations, we prove a local constancy result for reductions modulo prime powers of trianguline representations of generic dimension $d$. Moreover, we present two explicit applications. First, in dimension two, we extend to a prime power setting and to the whole rigid projective line a recent result of Bergdall, Levin and Liu concerning reductions of semi-stable representations of $\text{Gal}(\overline{\mathbb{Q}}_p / \mathbb{Q}_p)$ with fixed Hodge-Tate weights and large $\mathcal{L}$-invariant. Second, in dimension $d$, let $V_n$ be a sequence of crystalline representations converging in a certain geometric sense to a crystalline representation $V$. We show that for any refined version $(V, σ)$ of $V$ (or equivalently for any chosen triangulation of its attached $(\varphi, Γ)$-module $D_{\text{rig}} (V)$ over the Robba ring), there exists a sequence of refinement $σ_n$ of each of the $V_n$ such that the limit as refined representations $(V_n , σ_n )$ converges to the $(V, σ)$. This result does not hold under the weaker assumption that $V_n$ converges only uniformly $p$-adically to $V$ (in the sense of Chenevier, Khare and Larsen).2023-12-31T11:37:24ZEmiliano Tortihttp://arxiv.org/abs/2401.00404v2Explicit Generators for the Stabilizers of Rational Points in Thompson's Group $F$2024-11-20T05:39:38ZWe construct explicit finite generating sets for the stabilizers in Thompson's group $F$ of rational points of a unit interval or a Cantor set. Our technique is based on the Reidemeister-Schreier procedure in the context of Schreier graphs of such stabilizers in $F$. It is well known that the stabilizers of dyadic rational points are isomorphic to $F\times F$ and can thus be generated by 4 explicit elements. We show that the stabilizer of every non-dyadic rational point $b\in (0,1)$ is generated by 5 elements that are explicitly calculated as words in generators $x_0, x_1$ of $F$ that depend on the binary expansion of $b$. We also provide an alternative simple proof that the stabilizers of all rational points are finitely presented.2023-12-31T05:38:29Z19 pages, 9 figures and picturesKrystofer BakerDmytro Savchukhttp://arxiv.org/abs/2401.00408v2Computing greatest common divisor of several parametric univariate polynomials via generalized subresultant polynomials2024-09-06T06:46:15ZIn this paper, we tackle the following problem: compute the gcd for several univariate polynomials with parametric coefficients. It amounts to partitioning the parameter space into ``cells'' so that the gcd has a uniform expression over each cell and constructing a uniform expression of gcd in each cell. We tackle the problem as follows. We begin by making a natural and obvious extension of subresultant polynomials of two polynomials to several polynomials. Then we develop the following structural theories about them.
1. We generalize Sylvester's theory to several polynomials, in order to obtain an elegant relationship between generalized subresultant polynomials and the gcd of several polynomials, yielding an elegant algorithm.
2. We generalize Habicht's theory to several polynomials, in order to obtain a systematic relationship between generalized subresultant polynomials and pseudo-remainders, yielding an efficient algorithm.
Using the generalized theories, we present a simple (structurally elegant) algorithm which is significantly more efficient (both in the output size and computing time) than algorithms based on previous approaches.2023-12-31T06:32:54ZHoon HongJing Yanghttp://arxiv.org/abs/2401.00403v2Overcome Modal Bias in Multi-modal Federated Learning via Balanced Modality Selection2024-07-28T14:33:47ZSelecting proper clients to participate in each federated learning (FL) round is critical to effectively harness a broad range of distributed data. Existing client selection methods simply consider the mining of distributed uni-modal data, yet, their effectiveness may diminish in multi-modal FL (MFL) as the modality imbalance problem not only impedes the collaborative local training but also leads to a severe global modality-level bias. We empirically reveal that local training with a certain single modality may contribute more to the global model than training with all local modalities. To effectively exploit the distributed multiple modalities, we propose a novel Balanced Modality Selection framework for MFL (BMSFed) to overcome the modal bias. On the one hand, we introduce a modal enhancement loss during local training to alleviate local imbalance based on the aggregated global prototypes. On the other hand, we propose the modality selection aiming to select subsets of local modalities with great diversity and achieving global modal balance simultaneously. Our extensive experiments on audio-visual, colored-gray, and front-back datasets showcase the superiority of BMSFed over baselines and its effectiveness in multi-modal data exploitation.2023-12-31T05:37:27ZAccepted by ECCV24, 23 pagesYunfeng FanWenchao XuHaozhao WangFushuo HuoJinyu ChenSong Guohttp://arxiv.org/abs/2401.00416v2SVFAP: Self-supervised Video Facial Affect Perceiver2024-10-01T07:55:22ZVideo-based facial affect analysis has recently attracted increasing attention owing to its critical role in human-computer interaction. Previous studies mainly focus on developing various deep learning architectures and training them in a fully supervised manner. Although significant progress has been achieved by these supervised methods, the longstanding lack of large-scale high-quality labeled data severely hinders their further improvements. Motivated by the recent success of self-supervised learning in computer vision, this paper introduces a self-supervised approach, termed Self-supervised Video Facial Affect Perceiver (SVFAP), to address the dilemma faced by supervised methods. Specifically, SVFAP leverages masked facial video autoencoding to perform self-supervised pre-training on massive unlabeled facial videos. Considering that large spatiotemporal redundancy exists in facial videos, we propose a novel temporal pyramid and spatial bottleneck Transformer as the encoder of SVFAP, which not only largely reduces computational costs but also achieves excellent performance. To verify the effectiveness of our method, we conduct experiments on nine datasets spanning three downstream tasks, including dynamic facial expression recognition, dimensional emotion recognition, and personality recognition. Comprehensive results demonstrate that SVFAP can learn powerful affect-related representations via large-scale self-supervised pre-training and it significantly outperforms previous state-of-the-art methods on all datasets. Code is available at https://github.com/sunlicai/SVFAP.2023-12-31T07:44:05ZPublished in: IEEE Transactions on Affective Computing (Early Access). The code and models are available at https://github.com/sunlicai/SVFAPIEEE Transactions on Affective Computing, 2024Licai SunZheng LianKexin WangYu HeMingyu XuHaiyang SunBin LiuJianhua Tao10.1109/TAFFC.2024.3436913http://arxiv.org/abs/2401.00450v1Fault-tolerant quantum computation by hybrid qubits with bosonic cat-code and single photons2023-12-31T10:57:31ZHybridizing different degrees of freedom or physical platforms potentially offers various advantages in building scalable quantum architectures. We here introduce a fault-tolerant hybrid quantum computation by taking the advantages of both discrete variable (DV) and continuous variable (CV) systems. Particularly, we define a CV-DV hybrid qubit with bosonic cat-code and single photon, which is implementable in current photonic platforms. By the cat-code encoded in the CV part, the dominant loss errors are readily correctable without multi-qubit encoding, while the logical basis is inherently orthogonal due to the DV part. We design fault-tolerant architectures by concatenating hybrid qubits and an outer DV quantum error correction code such as topological codes, exploring their potential merits in developing scalable quantum computation. We demonstrate by numerical simulations that our scheme is at least an order of magnitude more resource-efficient over all previous proposals in photonic platforms, allowing to achieve a record-high loss threshold among existing CV and hybrid approaches. We discuss its realization not only in all-photonic platforms but also in other hybrid platforms including superconduting and trapped-ion systems, which allows us to find various efficient routes towards fault-tolerant quantum computing.2023-12-31T10:57:31Z21 pages, 8 figuresPRX Quantum 5, 030322 (2024)Jaehak LeeNuri KangSeok-Hyung LeeHyunseok JeongLiang JiangSeung-Woo Lee10.1103/PRXQuantum.5.030322http://arxiv.org/abs/2401.00469v1Exploring the Synergy: A Review of Dual-Functional Radar Communication Systems2023-12-31T11:55:09ZThis review paper examines the concept and advancements in the evolving landscape of Dual-functional Radar Communication (DFRC) systems. Traditionally, radar and communication systems have functioned independently, but current research is actively investigating the integration of these functionalities into a unified platform. This paper discusses the motivations behind the development of DFRC systems, the challenges involved, and the potential benefits they offer. A discussion on the performance bounds for DFRC systems is also presented. The paper encompasses a comprehensive analysis of various techniques, architectures, and technologies used in the design and optimization of DFRC systems, along with their performance and trade-offs. Additionally, we explore potential application scenarios for these joint communication and sensing systems, offering a comprehensive perspective on the multifaceted landscape of DFRC technology.2023-12-31T11:55:09Z17 pages, 7 figuresIEEE Aerospace and Electronic Systems Magazine, 41, 2026, 94-126Ali HanifSajid AhmedTareq Y. Al-NaffouriMohamed-Slim Alouin10.1109/MAES.2025.3551690http://arxiv.org/abs/2401.00405v1Generalizing Single-View 3D Shape Retrieval to Occlusions and Unseen Objects2023-12-31T05:39:38ZSingle-view 3D shape retrieval is a challenging task that is increasingly important with the growth of available 3D data. Prior work that has studied this task has not focused on evaluating how realistic occlusions impact performance, and how shape retrieval methods generalize to scenarios where either the target 3D shape database contains unseen shapes, or the input image contains unseen objects. In this paper, we systematically evaluate single-view 3D shape retrieval along three different axes: the presence of object occlusions and truncations, generalization to unseen 3D shape data, and generalization to unseen objects in the input images. We standardize two existing datasets of real images and propose a dataset generation pipeline to produce a synthetic dataset of scenes with multiple objects exhibiting realistic occlusions. Our experiments show that training on occlusion-free data as was commonly done in prior work leads to significant performance degradation for inputs with occlusion. We find that that by first pretraining on our synthetic dataset with occlusions and then finetuning on real data, we can significantly outperform models from prior work and demonstrate robustness to both unseen 3D shapes and unseen objects.2023-12-31T05:39:38ZQirui WuDaniel RitchieManolis SavvaAngel X. Changhttp://arxiv.org/abs/2401.00426v1keqing: knowledge-based question answering is a nature chain-of-thought mentor of LLM2023-12-31T08:39:04ZLarge language models (LLMs) have exhibited remarkable performance on various natural language processing (NLP) tasks, especially for question answering. However, in the face of problems beyond the scope of knowledge, these LLMs tend to talk nonsense with a straight face, where the potential solution could be incorporating an Information Retrieval (IR) module and generating response based on these retrieved knowledge. In this paper, we present a novel framework to assist LLMs, such as ChatGPT, to retrieve question-related structured information on the knowledge graph, and demonstrate that Knowledge-based question answering (Keqing) could be a nature Chain-of-Thought (CoT) mentor to guide the LLM to sequentially find the answer entities of a complex question through interpretable logical chains. Specifically, the workflow of Keqing will execute decomposing a complex question according to predefined templates, retrieving candidate entities on knowledge graph, reasoning answers of sub-questions, and finally generating response with reasoning paths, which greatly improves the reliability of LLM's response. The experimental results on KBQA datasets show that Keqing can achieve competitive performance and illustrate the logic of answering each question.2023-12-31T08:39:04Z12 pages, 6 figuresChaojie WangYishi XuZhong PengChenxi ZhangBo ChenXinrun WangLei FengBo Anhttp://arxiv.org/abs/2401.00437v1BatchEval: Towards Human-like Text Evaluation2023-12-31T09:34:51ZSignificant progress has been made in automatic text evaluation with the introduction of large language models (LLMs) as evaluators. However, current sample-wise evaluation paradigm suffers from the following issues: (1) Sensitive to prompt design; (2) Poor resistance to noise; (3) Inferior ensemble performance with static reference. Inspired by the fact that humans treat both criterion definition and inter sample comparison as references for evaluation, we propose BatchEval, a paradigm that conducts batch-wise evaluation iteratively to alleviate the above problems. We explore variants under this paradigm and confirm the optimal settings are two stage procedure with heterogeneous batch composition strategy and decimal scoring format. Comprehensive experiments across 3 LLMs on 4 text evaluation tasks demonstrate that BatchEval outperforms state-of-the-art methods by 10.5% on Pearson correlations with only 64% API cost on average. Further analyses have been conducted to verify the robustness, generalization, and working mechanism of BatchEval.2023-12-31T09:34:51Z19 pages, 9 figuresPeiwen YuanShaoxiong FengYiwei LiXinglin WangBoyuan PanHeda WangKan Lihttp://arxiv.org/abs/2401.00440v2TSGAN: An Optical-to-SAR Dual Conditional GAN for Optical based SAR Temporal Shifting2024-01-04T09:43:33ZIn contrast to the well-investigated field of SAR-to-Optical translation, this study explores the lesser-investigated domain of Optical-to-SAR translation, a challenging field due to the ill-posed nature of this translation. The complexity arises as a single optical data can have multiple SAR representations based on the SAR viewing geometry. We propose a novel approach, termed SAR Temporal Shifting, which inputs an optical data from the desired timestamp along with a SAR data from a different temporal point but with a consistent viewing geometry as the expected SAR data, both complemented with a change map of optical data during the intervening period. This model modifies the SAR data based on the changes observed in optical data to generate the SAR data for the desired timestamp. Our model, a dual conditional Generative Adversarial Network (GAN), named Temporal Shifting GAN (TSGAN), incorporates a siamese encoder in both the Generator and the Discriminator. To prevent the model from overfitting on the input SAR data, we employed a change weighted loss function. Our approach surpasses traditional translation methods by eliminating the GAN's fiction phenomenon, particularly in unchanged regions, resulting in higher SSIM and PSNR in these areas. Additionally, modifications to the Pix2Pix architecture and the inclusion of attention mechanisms have enhanced the model's performance on all regions of the data. This research paves the way for leveraging legacy optical datasets, the most abundant and longstanding source of Earth imagery data, extending their use to SAR domains and temporal analyses. To foster further research, we provide the code, datasets used in our study, and a framework for generating paired SAR-Optical datasets for new regions of interest. These resources are available on github.com/moienr/TemporalGAN2023-12-31T09:38:53ZComments: Added acknowledgments and corrected a typo. No changes to the main contentMoien RangzanSara AttarchiRichard GloaguenSeyed Kazem Alavipanahhttp://arxiv.org/abs/2401.00441v1Quantitative unique continuation for real-valued solutions to second order elliptic equations in the plane2023-12-31T09:45:09ZIn this article, we study a quantitative form of the Landis conjecture on exponential decay for real-valued solutions to second order elliptic equations with variable coefficients in the plane. In particular, we prove the following qualitative form of Landis conjecture, for $W_1, W_2 \in L^{\infty}(\mathbb R^2;\mathbb R^2)$, $V \in L^{\infty}(\mathbb R^2;\mathbb R)$ and $u \in H_{\mathrm{loc}}^{1}(\mathbb R^2)$ a real-valued weak solution to $-Δu - \nabla \cdot ( W_1 u ) +W_2 \cdot \nabla u + V u = 0$ in $\mathbb R^2$, satisfying for $δ>0$, $|u(x)| \leq \exp(- |x|^{1+δ})$, $x \in \mathbb R^2$, then $u \equiv 0$. Our methodology of proof is inspired by the one recently developed by Logunov, Malinnikova, Nadirashvili, and Nazarov that have treated the equation $-Δu + V u = 0$ in $\mathbb R^2$. Nevertheless, several differences and additional difficulties appear. New weak quantitative maximum principles are established for the construction of a positive multiplier in a suitable perforated domain, depending on the nodal set of $u$. The resulted divergence elliptic equation is then transformed into a non-homogeneous $\partial_{\overline{z}}$ equation thanks to a generalization of Stoilow factorization theorem obtained by the theory of quasiconformal mappings, an approximate type Poincaré lemma and the use of the Cauchy transform. Finally, a suitable Carleman estimate applied to the operator $\partial_{\overline{z}}$ is the last ingredient of our proof.2023-12-31T09:45:09ZComments welcomeKévin Le Balc'hDiego A. Souzahttp://arxiv.org/abs/2401.00466v1Online Symbolic Music Alignment with Offline Reinforcement Learning2023-12-31T11:42:42ZSymbolic Music Alignment is the process of matching performed MIDI notes to corresponding score notes. In this paper, we introduce a reinforcement learning (RL)-based online symbolic music alignment technique. The RL agent - an attention-based neural network - iteratively estimates the current score position from local score and performance contexts. For this symbolic alignment task, environment states can be sampled exhaustively and the reward is dense, rendering a formulation as a simplified offline RL problem straightforward. We evaluate the trained agent in three ways. First, in its capacity to identify correct score positions for sampled test contexts; second, as the core technique of a complete algorithm for symbolic online note-wise alignment; and finally, as a real-time symbolic score follower. We further investigate the pitch-based score and performance representations used as the agent's inputs. To this end, we develop a second model, a two-step Dynamic Time Warping (DTW)-based offline alignment algorithm leveraging the same input representation. The proposed model outperforms a state-of-the-art reference model of offline symbolic music alignment.2023-12-31T11:42:42ZProceedings of the 24th International Society for Music Information Retrieval Conference, {ISMIR} 2023, Milan, Italy, November 5-9, 2023Silvan David Peter10.5281/zenodo.10265367http://arxiv.org/abs/2401.00471v1Sounding Out Reconstruction Error-Based Evaluation of Generative Models of Expressive Performance2023-12-31T11:59:20ZGenerative models of expressive piano performance are usually assessed by comparing their predictions to a reference human performance. A generative algorithm is taken to be better than competing ones if it produces performances that are closer to a human reference performance. However, expert human performers can (and do) interpret music in different ways, making for different possible references, and quantitative closeness is not necessarily aligned with perceptual similarity, raising concerns about the validity of this evaluation approach. In this work, we present a number of experiments that shed light on this problem. Using precisely measured high-quality performances of classical piano music, we carry out a listening test indicating that listeners can sometimes perceive subtle performance difference that go unnoticed under quantitative evaluation. We further present tests that indicate that such evaluation frameworks show a lot of variability in reliability and validity across different reference performances and pieces. We discuss these results and their implications for quantitative evaluation, and hope to foster a critical appreciation of the uncertainties involved in quantitative assessments of such performances within the wider music information retrieval (MIR) community.2023-12-31T11:59:20Z10th International Conference on Digital Libraries for Musicology, November 10, 2023, Milan, ItalySilvan David PeterCarlos Eduardo Cancino-ChacónEmmanouil KarystinaiosGerhard Widmer10.1145/3625135.3625141http://arxiv.org/abs/2401.00479v1$L^p$ Maximal regularity for vector-valued Schrödinger operators2023-12-31T12:45:48ZIn this paper we consider the vector-valued Schrödinger operator $-Δ+ V$, where the potential term $V$ is a matrix-valued function whose entries belong to $L^1_{\rm loc}(\mathbb{R}^d)$ and, for every $x\in\mathbb{R}^d$, $V(x)$ is a symmetric and nonnegative definite matrix, with non positive off-diagonal terms and with eigenvalues comparable each other. For this class of potential terms we obtain maximal inequality in $L^1(\mathbb{R}^d,\mathbb{R}^m).$ Assuming further that the minimal eigenvalue of $V$ belongs to some reverse Hölder class of order $q\in(1,\infty)\cup\{\infty\}$, we obtain maximal inequality in $L^p(\mathbb{R}^d,\mathbb{R}^m)$, for $p$ in between $1$ and some $q$.2023-12-31T12:45:48ZDavide AddonaVincenzo LeoneLuca LorenziAbdelaziz Rhandihttp://arxiv.org/abs/2401.00487v1Spinterface Mediated Magnetic Properties of Co20Fe60B20/Alq3 Heterostructures2023-12-31T13:01:50ZOrganic semiconductors (OSCs) are suitable materials for spintronics applications as they form a spinterface when placed next to a ferromagnet, which in turn leads to novel functionalities. The evolution of spinterface can tune the global magnetic anisotropy, magnetization reversal, magnetization dynamics, etc. Planar tris-(8-hydroxyquinoline)aluminum (Alq3) OSC has shown tremendous potential for spintronics applications, thanks to its efficient spin-polarized current transport ability. Here, we establish the spinterface when the Alq3 molecules are deposited on amorphous ferromagnet Co20Fe60B20(CFB). The $π$-d hybridization in CFB/Alq3 enhances the coercive field and significantly modifies the shape and size of the magnetic domains. A $\sim$100% increase in uniaxial anisotropic energies and a reduction in magnetic damping are also evident owing to the strong interfacial hybridization.2023-12-31T13:01:50ZSwayang Priya MahantaAntarjami SahooSagarika NayakT. P. A. HaseDel AtkinsonSubhankar Bedantahttp://arxiv.org/abs/2401.00488v1The Sonified Hertzsprung-Russell Diagram2023-12-31T13:04:42ZUnderstanding the physical properties of stars, and putting these properties into the context of stellar evolution, is a core challenge in astronomical research. A key visualization in studying stellar evolution is the Hertzsprung-Russell diagram (HRD), organizing data about stellar luminosity and colour into a form that is informative about stellar structure and evolution. However, connecting the HRD with other sources of information, including stellar time series, is an outstanding challenge. Here we present a new method to turn stellar time series into sound. This method encodes physically meaningful features such that auditory comparisons between sonifications of different stars preserve astrophysical differences between them. We present an interactive multimedia version of the HRD that combines both visual and auditory components and that allows exploration of different types of stars both on and off the main sequence through both visual and auditory media.2023-12-31T13:04:42Z8 pages, 5 figures; accepted for publication in the proceedings of "The 28th International Conference on Auditory Display (ICAD 2023) - Special Session on Astronomical Data Sonification"Daniela HuppenkothenJuan PampinJames R. A. DavenportJames Wenlockhttp://arxiv.org/abs/2401.00449v1Teaching Digital Accessibility to Industry Professionals using the Community of Practice Framework: An Experience Report2023-12-31T10:55:26ZDespite recent initiatives aimed at improving accessibility, the field of digital accessibility remains markedly behind contemporary advancements in the software industry as a large number of real world software and web applications continue to fall short of accessibility requirements. A persisting skills deficit within the existing technology workforce has been an enduring impediment, hindering organizations from delivering truly accessible software products. This, in turn, elevates the risk of isolating and excluding a substantial portion of potential users. In this paper, we report lessons learned from a training program for teaching digital accessibility using the Communities of Practice (CoP) framework to industry professionals. We recruited 66 participants from a large multi-national software company and assigned them to two groups: one participating in a CoP and the other using self-paced learning. We report experiences from designing the training program, conducting the actual training, and assessing the efficiency of the two approaches. Based on these findings, we provide recommendations for practitioners in Learning and Development teams and educators in designing accessibility courses for industry professionals.2023-12-31T10:55:26ZTo be published in International Conference on Software Engineering (ICSE'24), Software Engineering Education and Training TrackParthasarathy PDSwaroop Joshi10.1145/3639474.3640083http://arxiv.org/abs/2401.00443v2Data-driven Energy Efficiency Modelling in Large-scale Networks: An Expert Knowledge and ML-based Approach2024-06-04T10:01:55ZThe energy consumption of mobile networks poses a critical challenge. Mitigating this concern necessitates the deployment and optimization of network energy-saving solutions, such as carrier shutdown, to dynamically manage network resources. Traditional optimization approaches encounter complexity due to factors like the large number of cells, stochastic traffic, channel variations, and intricate trade-offs. This paper introduces the simulated reality of communication networks (SRCON) framework, a novel, data-driven modeling paradigm that harnesses live network data and employs a blend of machine learning (ML)- and expert-based models. These mix of models accurately characterizes the functioning of network components, and predicts network energy efficiency and user equipment (UE) quality of service for any energy carrier shutdown configuration in a specific network. Distinguishing itself from existing methods, SRCON eliminates the reliance on expensive expert knowledge, drive testing, or incomplete maps for predicting network performance. This paper details the pipeline employed by SRCON to decompose the large network energy efficiency modeling problem into ML and expert-based submodels. It demonstrates how, by embracing stochasticity, and carefully crafting the relationship between such submodels, the overall computational complexity can be reduced and prediction accuracy enhanced. Results derived from real network data underscore the paradigm shift introduced by SRCON, showcasing significant gains over a state-of-the art method used by a operator for network energy efficiency modeling. The reliability of this local, data-driven modeling of the network proves to be a key asset for network energy-saving optimization.2023-12-31T10:03:08Z24 pages, 13 figures, submitted to IEEE Transactions on Machine Learning in Communications and NetworkingDavid López-PérezAntonio De DomenicoNicola PiovesanMerouane Debbahhttp://arxiv.org/abs/2401.00442v2A Comprehensive Overview of Fish-Eye Camera Distortion Correction Methods2024-05-13T15:18:57ZThe fisheye camera, with its unique wide field of view and other characteristics, has found extensive applications in various fields. However, the fisheye camera suffers from significant distortion compared to pinhole cameras, resulting in distorted images of captured objects. Fish-eye camera distortion is a common issue in digital image processing, requiring effective correction techniques to enhance image quality. This review provides a comprehensive overview of various methods used for fish-eye camera distortion correction. The article explores the polynomial distortion model, which utilizes polynomial functions to model and correct radial distortions. Additionally, alternative approaches such as panorama mapping, grid mapping, direct methods, and deep learning-based methods are discussed. The review highlights the advantages, limitations, and recent advancements of each method, enabling readers to make informed decisions based on their specific needs.2023-12-31T09:49:37ZJian XuDe-Wei HanKang LiJun-Jie LiZhao-Yuan Mahttp://arxiv.org/abs/2401.00425v2Geometric BV for twisted Courant sigma models and the BRST power finesse2024-11-05T16:02:51ZWe study twisted Courant sigma models, a class of topological field theories arising from the coupling of 3D 0-/2-form BF theory and Chern-Simons theory and containing a 4-form Wess-Zumino term. They are examples of theories featuring a nonlinearly open gauge algebra, where products of field equations appear in the commutator of gauge transformations, and they are reducible gauge systems. We determine the solution to the master equation using a technique, the BRST power finesse, that combines aspects of the AKSZ construction (which applies to the untwisted model) and the general BV-BRST formalism. This allows for a geometric interpretation of the BV coefficients in the interaction terms of the master action in terms of an induced generalised connection on a 4-form twisted (pre-)Courant algebroid, its Gualtieri torsion and the basic curvature tensor. It also produces a frame independent formulation of the model. We show, moreover, that the gauge fixed action is the sum of the classical one and a BRST commutator, as expected from a Schwarz type topological field theory.2023-12-31T08:38:23Z50 pages; published version with added references and minor corrections and improvementsJHEP 07 (2024) 115Athanasios ChatzistavrakidisNoriaki IkedaLarisa Jonke10.1007/JHEP07(2024)115http://arxiv.org/abs/2401.00411v3Study the structure of X(3872) from its lineshape2025-04-09T08:15:53ZWe fit the invariant mass distribution of ${X(3872)}\rightarrow{J}/ψπ^+π^-$ from LHCb using the propagator for S-wave near-threshold states in effective field theory. In this way, we can directly determine the $Z$ which measures the projection of the bound state on the compact state in ${X(3872)}$. Consequently, the structure of ${X(3872)}$ can be elucidated. Moreover, the fitting result also can describe well the data for ${X(3872)}\rightarrow{D}^{0}\overline{D}^{0*}$ from Belle experiment, which demonstrate the reliability of our fitting. The fitting indicates that $Z$ is a non-vanishing value within error, which supports that $X(3872)$ has a compact short-distant core.2023-12-31T06:54:42Z6 pages, 2 figures, 1 tableHongge XuNing YuZuman Zhanghttp://arxiv.org/abs/2401.00483v2Interacting ground states of moiré ladders2026-03-18T13:19:19ZMoiré materials have emerged as a rich platform for exploring strong correlation effects in low dimensions, with twisted bilayer graphene (TBG) as a paradigmatic example. To distill the essential ingredients driving moiré-induced phases, a simplified one-dimensional analog -- a two-leg ladder with spatially modulated interleg hopping and a uniform magnetic flux -- was recently introduced. This model, which we refer to as the moiré ladder, features a nearly flat lowest-energy band in a suitable parameter regime, capturing the band-flattening mechanism of TBG. We investigate the ground-state phase diagram of the moiré ladder using a combination of bosonization and density matrix renormalization group (DMRG) techniques, and systematically disentangle the respective roles of the flux and the hopping modulation. At half filling, previous numerical work identified a metal-insulator transition at finite interaction strength and an unexpected ferromagnetic ground state. Revisiting this, we show that the metal-insulator transition can be understood perturbatively within bosonization, governed by the number of Fermi points. In contrast, the ferromagnetic correlations are nonperturbative and require both flux and spatial modulation -- neither alone is sufficient. We extend our analysis to other fillings: one-quarter, three-quarters, slightly above half filling (half filling plus two electrons), and slightly below half filling (half filling minus two electrons). At moderate interactions, we observe ferromagnetism below half filling and antiferromagnetism above; at stronger interactions, ferromagnetism dominates across all studied fillings. Crucially, the analysis demonstrates that periodic interleg hopping alone does not engender new correlated phases; the magnetic flux is essential for the observed unconventional behavior.2023-12-31T12:56:09ZPhys. Rev. B 113, 104432 (2026)Paban Kumar PatraRanjith R. KumarYixuan HuangHridis K. Pal10.1103/8lz4-s1f9