https://arxiv.org/api/hdAEMxLRbPpWWcUVsXt8+OcubLIarXiv Query: search_query=&id_list=2401.00501,2401.00502,2401.00503,2401.00504,2401.00505,2401.00506,2401.00507,2401.00508,2401.00509,2401.00510,2401.00511,2401.00512,2401.00513,2401.00514,2401.00515,2401.00516,2401.00517,2401.00518,2401.00519,2401.00520,2401.00521,2401.00522,2401.00523,2401.00524,2401.00525,2401.00526,2401.00527,2401.00528,2401.00529,2401.00530,2401.00531,2401.00532,2401.00533,2401.00534,2401.00535,2401.00536,2401.00537,2401.00538,2401.00539,2401.00540,2401.00541,2401.00542,2401.00543,2401.00544,2401.00545,2401.00546,2401.00547,2401.00548,2401.00549,2401.00550,2401.00551,2401.00552,2401.00553,2401.00554,2401.00555,2401.00556,2401.00557,2401.00558,2401.00559,2401.00560,2401.00561,2401.00562,2401.00563,2401.00564,2401.00565,2401.00566,2401.00567,2401.00568,2401.00569,2401.00570,2401.00571,2401.00572,2401.00573,2401.00574,2401.00575,2401.00576,2401.00577,2401.00578,2401.00579,2401.00580,2401.00581,2401.00582,2401.00583,2401.00584,2401.00585,2401.00586,2401.00587,2401.00588,2401.00589,2401.00590,2401.00591,2401.00592,2401.00593,2401.00594,2401.00595,2401.00596,2401.00597,2401.00598,2401.00599,2401.00600&start=0&max_results=1002026-07-02T22:41:42Z1001000http://arxiv.org/abs/2401.00524v1Effect of Optimizer, Initializer, and Architecture of Hypernetworks on Continual Learning from Demonstration2023-12-31T15:43:09ZIn continual learning from demonstration (CLfD), a robot learns a sequence of real-world motion skills continually from human demonstrations. Recently, hypernetworks have been successful in solving this problem. In this paper, we perform an exploratory study of the effects of different optimizers, initializers, and network architectures on the continual learning performance of hypernetworks for CLfD. Our results show that adaptive learning rate optimizers work well, but initializers specially designed for hypernetworks offer no advantages for CLfD. We also show that hypernetworks that are capable of stable trajectory predictions are robust to different network architectures. Our open-source code is available at https://github.com/sebastianbergner/ExploringCLFD.2023-12-31T15:43:09ZSayantan AuddySebastian BergnerJustus Piaterhttp://arxiv.org/abs/2401.00536v1A Multi-Task, Multi-Modal Approach for Predicting Categorical and Dimensional Emotions2023-12-31T16:48:03ZSpeech emotion recognition (SER) has received a great deal of attention in recent years in the context of spontaneous conversations. While there have been notable results on datasets like the well known corpus of naturalistic dyadic conversations, IEMOCAP, for both the case of categorical and dimensional emotions, there are few papers which try to predict both paradigms at the same time. Therefore, in this work, we aim to highlight the performance contribution of multi-task learning by proposing a multi-task, multi-modal system that predicts categorical and dimensional emotions. The results emphasise the importance of cross-regularisation between the two types of emotions. Our approach consists of a multi-task, multi-modal architecture that uses parallel feature refinement through self-attention for the feature of each modality. In order to fuse the features, our model introduces a set of learnable bridge tokens that merge the acoustic and linguistic features with the help of cross-attention. Our experiments for categorical emotions on 10-fold validation yield results comparable to the current state-of-the-art. In our configuration, our multi-task approach provides better results compared to learning each paradigm separately. On top of that, our best performing model achieves a high result for valence compared to the previous multi-task experiments.2023-12-31T16:48:03ZCompanion Publication of the 25th International Conference on Multimodal Interaction (pp. 311-317)Alex-Răzvan IspasThéo Deschamps-BergerLaurence Devillers10.1145/3610661.3616190http://arxiv.org/abs/2401.00558v1Symmetrical Sonin kernels in terms of the hypergeometric functions2023-12-31T18:11:18ZIn this paper, we introduce a new class of the kernels of the integral transforms of the Laplace convolution type that we call symmetrical Sonin kernels. For a symmetrical Sonin kernel given in terms of some elementary or special functions, its associated kernel has the same form with possibly different parameter values. Several known and new kernels of this type are derived by means of the Sonin method in the time domain and using the Laplace integral transform in the frequency domain. The new symmetrical Sonin kernels are provided in terms of the Wright function and some extensions of the Horn confluent hypergeometric functions in two variables.2023-12-31T18:11:18Z23 pagesYuri Luchkohttp://arxiv.org/abs/2401.00564v1Time-dependent backgrounds from marginal deformations of Minimal Strings in AdS$_3$2023-12-31T18:51:14ZWe study a class of time-dependent backgrounds in string theory which consist of marginal deformations of minimal strings on AdS$_3$. For such backgrounds, we compute the three-point amplitudes and analyze their properties.2023-12-31T18:51:14Z18 pagesEoin DowdGaston Giribethttp://arxiv.org/abs/2401.00572v1Sealed Kurepa Trees2023-12-31T19:19:17ZIn this paper we investigate the problem of the distributivity of Kurepa trees. We show that it is consistent that there are Kurepa trees and for every Kurepa tree there is a small forcing notion which adds a branch to it without collapsing cardinals. On the other hand, we derive a proper forcing notion for making an arbitrary Kurepa tree into a non-distributive tree without collapsing $\aleph_1$ and $\aleph_2$.2023-12-31T19:19:17ZItamar GironYair Hayuthttp://arxiv.org/abs/2401.00580v1Multimodal surface coils for low-field MR imaging2023-12-31T20:04:27ZLeveraging the potential of low-field Magnetic Resonance Imaging (MRI), our study introduces the multimodal surface RF coil, a design tailored to overcome the limitations of conventional coils in this context. The inherent challenges of low-field MRI, notably suboptimal signal-to-noise ratio (SNR) and the need for specialized RF coils, are effectively addressed by our novel design. The multimodal surface coil is characterized by a unique assembly of resonators, optimized for both B1 efficiency and low-frequency tuning capabilities, essential for low-field applications. This paper provides a thorough investigation of the conceptual framework, design intricacies, and bench test validation of the multimodal surface coil. Through detailed simulations and comparative analyses, we demonstrate its superior performance in terms of B1 field efficiency, outperforming conventional surface coils.2023-12-31T20:04:27ZYunkun ZhaoAditya A BhosaleXiaoliang Zhanghttp://arxiv.org/abs/2401.00582v1An Analysis of Embedding Layers and Similarity Scores using Siamese Neural Networks2023-12-31T20:21:58ZLarge Lanugage Models (LLMs) are gaining increasing popularity in a variety of use cases, from language understanding and writing to assistance in application development. One of the most important aspects for optimal funcionality of LLMs is embedding layers. Word embeddings are distributed representations of words in a continuous vector space. In the context of LLMs, words or tokens from the input text are transformed into high-dimensional vectors using unique algorithms specific to the model. Our research examines the embedding algorithms from leading companies in the industry, such as OpenAI, Google's PaLM, and BERT. Using medical data, we have analyzed similarity scores of each embedding layer, observing differences in performance among each algorithm. To enhance each model and provide an additional encoding layer, we also implemented Siamese Neural Networks. After observing changes in performance with the addition of the model, we measured the carbon footage per epoch of training. The carbon footprint associated with large language models (LLMs) is a significant concern, and should be taken into consideration when selecting algorithms for a variety of use cases. Overall, our research compared the accuracy different, leading embedding algorithms and their carbon footage, allowing for a holistic review of each embedding algorithm.2023-12-31T20:21:58Z10 pages, 11 figuresYash BingiYiqiao Yinhttp://arxiv.org/abs/2401.00502v3On the dilation current in metric-affine gravity2024-04-16T07:05:41ZWe review $F(R,\mathcal{D})$ gravity in the metric-affine framework, where $\mathcal{D}$ is the divergence of the dilation current appearing in the hypermomentum tensor. We assume only linear couplings between the general affine connection and the matter fields (minimal coupling) and break projective invariance to preserve a nonvanishing dilation current. For $F(R,\mathcal{D})$ linear in $\mathcal{D}$ the dilation current dependence in the function $F(R,\mathcal{D})$ does not contribute to the field equations of the theory.
We show that, on the other hand, in more complicated cases (e.g., considering the function $F(R,\mathcal{D})=R+α\mathcal{D}^2$), the $\mathcal{D}$ contribution to the metric field equations is nontrivial and can affect the cosmology of the theory.2023-12-31T13:49:31Z17 pages, comments and conclusions added, misprints corrected, accepted for publication in Annals of PhysicsAnnals Phys. 465 (2024) 169664D. KenzhalinS. MyrzakulR. MyrzakulovL. Ravera10.1016/j.aop.2024.169664http://arxiv.org/abs/2401.00531v3Contractibility of the orbit space of a saturated fusion system after Steinberg2024-10-15T23:22:19ZRecently, Steinberg used discrete Morse theory to give a new proof of a theorem of Symonds that the orbit space of the poset of nontrivial $p$-subgroups of a finite group is contractible. We extend Steinberg's argument in two ways, covering more general versions of the theorem that were already known. In particular, following a strategy of Libman, we give a discrete Morse theoretic argument for the contractibility of the orbit space of a saturated fusion system.2023-12-31T16:22:37Zv2: Revisions in response to a referee report, including correction of an error in the construction of the Morse matchingOmar DennaouiJonathon Villarealhttp://arxiv.org/abs/2401.00561v2QGLAB: A MATLAB Package for Computations on Quantum Graphs2024-11-07T18:36:00ZWe describe QGLAB, a new MATLAB package for analyzing partial differential equations on quantum graphs. The software is built on the existing, object-oriented MATLAB directed-graph class, inheriting its structure and adding additional easy-to-use features. The package allows one to construct a quantum graph and accurately compute the spectrum of elliptic operators, solutions to Poisson problems, the linear and nonlinear time evolution of a variety of PDEs, the continuation of branches of steady states (including locating and switching branches at bifurcations) and more. It overcomes the major challenge of discretizing quantum graphs -- the enforcement of vertex conditions -- using non-square differentiation matrices. It uses a unified framework to implement finite-difference and Chebyshev discretizations of differential operators on a quantum graph. For simplicity, the package overloads many built-in MATLAB functions to work on the class.2023-12-31T18:23:28Z41 pages, 18 figures. Major rewrite. Examples moved from appendix to body. Comments Welcome! Code associated with this publication available at https://github.com/manroygood/Quantum-Graphs/tree/masterRoy H. GoodmanGrace ConteJeremy L. Marzuolahttp://arxiv.org/abs/2401.00523v2Compressing Deep Image Super-resolution Models2024-02-21T20:25:53ZDeep learning techniques have been applied in the context of image super-resolution (SR), achieving remarkable advances in terms of reconstruction performance. Existing techniques typically employ highly complex model structures which result in large model sizes and slow inference speeds. This often leads to high energy consumption and restricts their adoption for practical applications. To address this issue, this work employs a three-stage workflow for compressing deep SR models which significantly reduces their memory requirement. Restoration performance has been maintained through teacher-student knowledge distillation using a newly designed distillation loss. We have applied this approach to two popular image super-resolution networks, SwinIR and EDSR, to demonstrate its effectiveness. The resulting compact models, SwinIRmini and EDSRmini, attain an 89% and 96% reduction in both model size and floating-point operations (FLOPs) respectively, compared to their original versions. They also retain competitive super-resolution performance compared to their original models and other commonly used SR approaches. The source code and pre-trained models for these two lightweight SR approaches are released at https://pikapi22.github.io/CDISM/.2023-12-31T15:38:50ZYuxuan JiangJakub NawalaFan ZhangDavid Bull10.1109/PCS60826.2024.10566374http://arxiv.org/abs/2401.00508v1Amplification of quantum transfer and quantum ratchet2023-12-31T14:04:43ZAmplification of quantum transfer and ratchet--type processes are important for quantum technologies. We also expect that quantum ratchet works in quantum photosynthesis, where possible role of quantum effects is now widely discussed but the underlying dynamical processes are still not clearly known. In this work, we study a model of amplification of quantum transfer and making it directed which we call the quantum ratchet model. The model is based on a special quantum control master equation with dynamics induced by a feedback-type process. The ratchet effect is achieved in the quantum control model with dissipation and sink, where the Hamiltonian depends on vibrations in the energy difference synchronized with transitions between energy levels. A similarity between this model and the model of coherent transport in quantum photosynthesis, where the time dependence of the Hamiltonian arises due to vibrons, is studied. Amplitude and frequency of the oscillating vibron together with the dephasing rate are the parameters of the quantum ratchet which determine its efficiency. We study with which parameters the quantum ratchet minimizes the exction recombination time and show that the experimentally known values of the parameters of the photosynthetic reaction center correspond to values of the parameters of the quantum ratchet which realize a local minimum of the exciton recombination time. We also find another values of the parameters of the quantum ratchet minimizing the exciton recombination time, which corresponds to a twice smaller frequency of the vibron compared to that observed in experiments.2023-12-31T14:04:43Z6 figuresPhysica Scripta, Vol. 98, No 12, 125122 (2023)Sergei KozyrevAlexander Pechen10.1088/1402-4896/ad0c3dhttp://arxiv.org/abs/2401.00503v1Viz: A QLoRA-based Copyright Marketplace for Legally Compliant Generative AI2023-12-31T13:53:06ZThis paper aims to introduce and analyze the Viz system in a comprehensive way, a novel system architecture that integrates Quantized Low-Rank Adapters (QLoRA) to fine-tune large language models (LLM) within a legally compliant and resource efficient marketplace. Viz represents a significant contribution to the field of artificial intelligence, particularly in addressing the challenges of computational efficiency, legal compliance, and economic sustainability in the utilization and monetization of LLMs. The paper delineates the scholarly discourse and developments that have informed the creation of Viz, focusing primarily on the advancements in LLM models, copyright issues in AI training (NYT case, 2023), and the evolution of model fine-tuning techniques, particularly low-rank adapters and quantized low-rank adapters, to create a sustainable and economically compliant framework for LLM utilization. The economic model it proposes benefits content creators, AI developers, and end-users, delineating a harmonious integration of technology, economy, and law, offering a comprehensive solution to the complex challenges of today's AI landscape.2023-12-31T13:53:06ZDipankar Sarkarhttp://arxiv.org/abs/2401.00504v1HSC-GPT: A Large Language Model for Human Settlements Construction2023-12-31T13:56:15ZThe field of human settlement construction encompasses a range of spatial designs and management tasks, including urban planning and landscape architecture design. These tasks involve a plethora of instructions and descriptions presented in natural language, which are essential for understanding design requirements and producing effective design solutions. Recent research has sought to integrate natural language processing (NLP) and generative artificial intelligence (AI) into human settlement construction tasks. Due to the efficient processing and analysis capabilities of AI with data, significant successes have been achieved in design within this domain. However, this task still faces several fundamental challenges. The semantic information involved includes complex spatial details, diverse data source formats, high sensitivity to regional culture, and demanding requirements for innovation and rigor in work scenarios. These factors lead to limitations when applying general generative AI in this field, further exacerbated by a lack of high-quality data for model training. To address these challenges, this paper first proposes HSC-GPT, a large-scale language model framework specifically designed for tasks in human settlement construction, considering the unique characteristics of this domain.2023-12-31T13:56:15ZChen RanYao XueqiJiang XuhuiHan ZhengqiGuo JingzeZhang XianyueLin ChunyuLiu ChuminZhao JingLian ZekeZhang JingjingLi Kekehttp://arxiv.org/abs/2401.00509v2Twisted products: Enveloping actions and equivariant absolute neighborhood extensors2024-01-03T14:07:43ZThe classical notion of twisted product is studied in the context of partial actions, in particular, we show that the globalization of a partial action is a twisted product. In addition, we establish conditions for the metrizability of twisted products, and some homotopy and categorical properties are proved. Furthermore, sufficient conditions for the enveloping space to be an equivariant absolute neighborhood extensor are also studied.2023-12-31T14:08:43ZLuis MartínezHéctor Pinedohttp://arxiv.org/abs/2401.00520v1Monte Carlo Expectation-Maximization algorithm to detect imprinting and maternal effects for discordant sib-pair data2023-12-31T15:17:02ZNumerous statistical methods have been developed to explore genomic imprinting and maternal effects, which are causes of parent-of-origin patterns in complex human diseases. Most of the methods, however, either only model one of these two confounded epigenetic effects, or make strong yet unrealistic assumptions about the population to avoid over-parameterization. A recent partial likelihood method (LIMEDSP ) can identify both epigenetic effects based on discordant sibpair family data without those assumptions. Theoretical and empirical studies have shown its validity and robustness. As LIMEDSP method obtains parameter estimation by maximizing partial likelihood, it is interesting to compare its efficiency with full likelihood maximizer. To overcome the difficulty in over-parameterization when using full likelihood, this study proposes a discordant sib-pair design based Monte Carlo Expectation Maximization (MCEMDSP ) method to detect imprinting and maternal effects jointly. Those unknown mating type probabilities, the nuisance parameters, are considered as latent variables in EM algorithm. Monte Carlo samples are used to numerically approximate the expectation function that cannot be solved algebraically. Our simulation results show that though this MCEMDSP algorithm takes longer computation time, it can generally detect both epigenetic effects with higher power, which demonstrates that it can be a good complement of LIMEDSP method2023-12-31T15:17:02ZRuwani HerathAlex TrindadeFangyuan Zhanghttp://arxiv.org/abs/2401.00521v1Multi-spatial Multi-temporal Air Quality Forecasting with Integrated Monitoring and Reanalysis Data2023-12-31T15:25:23ZAccurate air quality forecasting is crucial for public health, environmental monitoring and protection, and urban planning. However, existing methods fail to effectively utilize multi-scale information, both spatially and temporally. Spatially, there is a lack of integration between individual monitoring stations and city-wide scales. Temporally, the periodic nature of air quality variations is often overlooked or inadequately considered. To address these limitations, we present a novel Multi-spatial Multi-temporal air quality forecasting method based on Graph Convolutional Networks and Gated Recurrent Units (M2G2), bridging the gap in air quality forecasting across spatial and temporal scales. The proposed framework consists of two modules: Multi-scale Spatial GCN (MS-GCN) for spatial information fusion and Multi-scale Temporal GRU(MT-GRU) for temporal information integration. In the spatial dimension, the MS-GCN module employs a bidirectional learnable structure and a residual structure, enabling comprehensive information exchange between individual monitoring stations and the city-scale graph. Regarding the temporal dimension, the MT-GRU module adaptively combines information from different temporal scales through parallel hidden states. Leveraging meteorological indicators and four air quality indicators, we present comprehensive comparative analyses and ablation experiments, showcasing the higher accuracy of M2G2 in comparison to nine currently available advanced approaches across all aspects. The improvements of M2G2 over the second-best method on RMSE of the 24h/48h/72h are as follows: PM2.5: (7.72%, 6.67%, 10.45%); PM10: (6.43%, 5.68%, 7.73%); NO2: (5.07%, 7.76%, 16.60%); O3: (6.46%, 6.86%, 9.79%). Furthermore, we demonstrate the effectiveness of each module of M2G2 by ablation study.2023-12-31T15:25:23ZYuxiao HuQian LiXiaodan ShiJinyue YanYuntian Chenhttp://arxiv.org/abs/2401.00525v1Pack and Measure: An Effective Approach for Influence Propagation in Social Networks2023-12-31T15:51:33ZThe Influence Maximization problem under the Independent Cascade model (IC) is considered. The problem asks for a minimal set of vertices to serve as "seed set" from which a maximum influence propagation is expected. New seed-set selection methods are introduced based on the notions of a $d$-packing and vertex centrality. In particular, we focus on selecting seed-vertices that are far apart and whose influence-values are the highest in their local communities. Our best results are achieved via an initial computation of a $d$-Packing followed by selecting either vertices of high degree or high centrality in their respective closed neighborhoods. This overall "Pack and Measure" approach proves highly effective as a seed selection method.2023-12-31T15:51:33ZFaisal N. Abu-KhzamGhinwa Bou MatarSergio Thoumihttp://arxiv.org/abs/2401.00530v1Probing topological phase transition with non-Hermitian perturbations2023-12-31T16:19:42ZWe demonstrate that non-Hermitian perturbations can probe topological phase transitions and unambiguously detect non-Abelian zero modes. We show that under carefully designed non-Hermitian perturbations, the Loschmidt echo(LE) decays into 1/N where N is the ground state degeneracy in the topological non-trivial phase, while it approaches 1 in the trivial phase. This distinction is robust against small parameter deviations in the non-Hermitian perturbations. We further study four well-known models that support Majorana or parafermionic zero modes. By calculating their dynamical responses to specific non-Hermitian perturbations, we prove that the steady-state LE can indeed differentiate between different phases. This method avoids the ambiguity introduced by trivial zero-energy states and thus provides an alternative and promising way to demonstrate the emergence of topologically non-trivial phases. The experimental realizations of non-Hermitian perturbations are discussed.2023-12-31T16:19:42Z11 pages, 4 figuresJingcheng LiangChen FangJiangping Huhttp://arxiv.org/abs/2401.00540v1Study Duration Prediction for Clinical Trials with Time-to-Event Endpoints Using Mixture Distributions Accounting for Heterogeneous Population2023-12-31T17:06:26ZIn the era of precision medicine, more and more clinical trials are now driven or guided by biomarkers, which are patient characteristics objectively measured and evaluated as indicators of normal biological processes, pathogenic processes, or pharmacologic responses to therapeutic interventions. With the overarching objective to optimize and personalize disease management, biomarker-guided clinical trials increase the efficiency by appropriately utilizing prognostic or predictive biomarkers in the design. However, the efficiency gain is often not quantitatively compared to the traditional all-comers design, in which a faster enrollment rate is expected (e.g. due to no restriction to biomarker positive patients) potentially leading to a shorter duration. To accurately predict biomarker-guided trial duration, we propose a general framework using mixture distributions accounting for heterogeneous population. Extensive simulations are performed to evaluate the impact of heterogeneous population and the dynamics of biomarker characteristics and disease on the study duration. Several influential parameters including median survival time, enrollment rate, biomarker prevalence and effect size are identitied. Re-assessments of two publicly available trials are conducted to empirically validate the prediction accuracy and to demonstrate the practical utility. The R package \emph{detest} is developed to implement the proposed method and is publicly available on CRAN.2023-12-31T17:06:26ZHong ZhangJie PuShibing DengSatrajit RoychoudhuryHaitao ChuDouglas Robinsonhttp://arxiv.org/abs/2401.00543v1A binomial random multigraph2023-12-31T17:11:33ZFix a positive integer $n$, a real number $p\in (0,1]$, and a (perhaps random) hypergraph $\mathcal{H}$ on $[n]$. We introduce and investigate the following random multigraph model, which we denote $\mathbb{G}(n,p\, ; \,\mathcal{H})$: begin with an empty graph on $n$ vertices, which are labelled by the set $[n]$. For every $H\in \mathcal{H}$ choose, independently from previous choices, a doubleton from $H$, say $D = \{i,j\} \subset H$, uniformly at random and then introduce an edge between the vertices $i$ and $j$ in the graph with probability $p$, where each edge is introduced independently of all other edges.2023-12-31T17:11:33Z20 pages. Comments are welcomeChristos Pelekishttp://arxiv.org/abs/2401.00548v1Cesàro summability of Taylor series in higher order weighted Dirichlet type spaces2023-12-31T17:34:59ZFor a positive integer $m$ and a finite non-negative Borel measure $μ$ on the unit circle, we study the Hadamard multipliers of higher order weighted Dirichlet-type spaces $\mathcal H_{μ, m}$. We show that if $α>\frac{1}{2},$ then for any $f$ in $\mathcal H_{μ, m},$ the sequence of generalized Ces{à}ro sums $\{σ_n^α[f]\}$ converges to $f$. We further show that if $α=\frac{1}{2}$ then for the Dirac delta measure supported at any point on the unit circle, the previous statement breaks down for every positive integer $m$.2023-12-31T17:34:59Z14 pages, comments and suggestions are welcomeSoumitra GharaRajeev GuptaMd. Ramiz Rezahttp://arxiv.org/abs/2401.00579v1Exploring the Effectiveness of Instruction Tuning in Biomedical Language Processing2023-12-31T20:02:10ZLarge Language Models (LLMs), particularly those similar to ChatGPT, have significantly influenced the field of Natural Language Processing (NLP). While these models excel in general language tasks, their performance in domain-specific downstream tasks such as biomedical and clinical Named Entity Recognition (NER), Relation Extraction (RE), and Medical Natural Language Inference (NLI) is still evolving. In this context, our study investigates the potential of instruction tuning for biomedical language processing, applying this technique to two general LLMs of substantial scale. We present a comprehensive, instruction-based model trained on a dataset that consists of approximately $200,000$ instruction-focused samples. This dataset represents a carefully curated compilation of existing data, meticulously adapted and reformatted to align with the specific requirements of our instruction-based tasks. This initiative represents an important step in utilising such models to achieve results on par with specialised encoder-only models like BioBERT and BioClinicalBERT for various classical biomedical NLP tasks. Our work includes an analysis of the dataset's composition and its impact on model performance, providing insights into the intricacies of instruction tuning. By sharing our codes, models, and the distinctively assembled instruction-based dataset, we seek to encourage ongoing research and development in this area.2023-12-31T20:02:10ZOmid RohanianMohammadmahdi NouriborjiDavid A. Cliftonhttp://arxiv.org/abs/2401.00589v12-flavour $SU(2)$ gauge theory with exponential clover Wilson fermions2023-12-31T21:16:32ZComposite Higgs models are a class of models proposed to address the hierarchy and naturalness problems associated with the Standard Model fundamental scalar Higgs. $SU(2)$ with two fundamental flavours is a minimal model for the composite Higgs sector which is not yet ruled out by experimental data. We present lattice results for $SU(2)$ with two fundamental mass degenerate flavours. For the fermion action we use the new exponential clover Wilson fermion action, which offers $O(a)$ improvement. We discuss tuning the $c_{\mathrm{SW}}$ parameter through Schrödinger functional simulations, the scale setting of the ensembles using the Wilson gauge flow, and the low energy spectroscopy of the theory including the masses of the pseudoscalar isotriplet Goldstone bosons and the vector isotriplet.2023-12-31T21:16:32ZProceedings of The 40th International Symposium on Lattice Field Theory (Lattice 2023)Laurence Sebastian BowesVincent DrachPatrick FritzschAntonio RagoFernando Romero-Lopezhttp://arxiv.org/abs/2401.00590v2Giant Optical Anisotropy in 2D Metal-Organic Chalcogenates2024-04-03T17:32:41ZOptical anisotropy is a fundamental attribute of some crystalline materials and is quantified via birefringence. A birefringent crystal not only gives rise to asymmetrical light propagation but also attenuation along two distinct polarizations, a phenomenon called linear dichroism (LD). Two-dimensional (2D) layered materials with high in- and out-of-plane anisotropy have garnered interest in this regard. Mithrene, a 2D metal-organic chalcogenate (MOCHA) compound, exhibits strong excitonic resonances due to its naturally occurring multi-quantum well (MQW) structure and in-plane anisotropic response in the blue wavelength (~400-500 nm) regime. The MQW structure and the large refractive indices of mithrene allow the hybridization of the excitons with photons to form self-hybridized exciton-polaritons in mithrene crystals with appropriate thicknesses. Here, we report the giant birefringence (~1.01) and tunable in-plane anisotropic response of mithrene, which stem from its low symmetry crystal structure and unique excitonic properties. We show that the LD in mithrene can be tuned by leveraging the anisotropic exciton-polariton formation via the cavity coupling effect exhibiting giant in-plane LD (~77.1%) at room temperature. Our results indicate that mithrene is an ideal polaritonic birefringent material for polarization-sensitive nanophotonic applications in the short wavelength regime.2023-12-31T21:21:51ZBongjun ChoiKiyoung JoMahfujur RahamanAdam AlfieriJason LynchGreg K. PribilHyeongjun KohEric A. StachDeep Jariwalahttp://arxiv.org/abs/2401.00562v1Ruhr Hand Motion Catalog of Human Center-Out Transport Trajectories in 3D Task-Space Captured by a Redundant Measurement System2023-12-31T18:39:42ZNeurological conditions are a major source of movement disorders. Motion modelling and variability analysis have the potential to identify pathology but require profound data. We introduce a systematic dataset of 3D center-out task-space trajectories of human hand transport movements in a natural setting. The transport tasks of this study consist of grasping a cylindric object from a unified start position and transporting it to one of nine target locations in unconstrained operational space. The measurement procedure is automatized to record ten trials per target location. With that, the dataset consists of 90 movement trajectories for each hand of 31 participants without known movement disorders. The participants are aged between 21 and 78 years, covering a wide range. Data are recorded redundantly by both an optical tracking system and an IMU sensor. As opposed to the stationary capturing system, the IMU can be considered as a portable, low-cost and energy-efficient alternative to be implemented on embedded systems.2023-12-31T18:39:42ZTim SziburisSusanne BlexTobias GlasmachersIoannis Iossifidishttp://arxiv.org/abs/2401.00586v3Scale Invariant Scattering and Bernoulli Numbers2024-10-24T06:21:24ZNon-relativistic quantum mechanical scattering from an inverse square potential in two spatial dimensions leads to a novel representation of the Bernoulli numbers.2023-12-31T20:41:29ZSIGMA 20 (2024), 096, 4 pagesThomas L. Curtright10.3842/SIGMA.2024.096http://arxiv.org/abs/2401.00549v3Democratic actions with scalar fields: symmetric sigma models, supergravity actions and the effective theory of the type IIB superstring2024-09-02T19:07:10ZThe dualization of the scalar fields of a theory into (d-2)-form potentials preserving all the global symmetries is one of the main problems in the construction of democratic pseudoactions containing simultaneously all the original fields and their duals. We study this problem starting with the simplest cases and we show how it can be solved for scalars parametrizing Riemannian symmetric sigma-models as in maximal and half-maximal supergravities. Then, we use this result to write democratic pseudoactions for theories in which the scalars are non-minimally coupled to (p+1)-form potentials in any dimension. These results include a proposal of democratic pseudoaction for the generic bosonic sector of 4-dimensional maximal and half-maximal ungauged supergravities. Furthermore, we propose a democratic pseudoaction for the bosonic sector of N=2B,d=10 supergravity (the effective action of the type IIB superstring theory) containing two 0-, two 2-, one 4-, two 6- and three 8-forms which is manifestly invariant under global SL(2,R) transformations.2023-12-31T17:36:39ZMisprints corrected in text and some equations. Main results unchanged. Version accepted for publication in SciPostJose Juan Fernandez-MelgarejoGiacomo GiorgiCarmen Gomez-FayrenTomas OrtinMatteo Zattihttp://arxiv.org/abs/2401.00559v2Perfect matchings and loose Hamilton cycles in the semirandom hypergraph model2024-09-24T22:28:49ZWe study the 2-offer semirandom 3-uniform hypergraph model on $n$ vertices. At each step, we are presented with 2 uniformly random vertices. We choose any other vertex, thus creating a hyperedge of size 3. We show a strategy that constructs a perfect matching, and another that constructs a loose Hamilton cycle, both succeeding asymptotically almost surely within $Θ(n)$ steps. Both results extend to $s$-uniform hypergraphs. The challenges with hypergraphs, and our methods, are qualitatively different from what has been seen for semirandom graphs. Much of our analysis is done on an auxiliary graph that is a uniform $k$-out subgraph of a random bipartite graph, and this tool may be useful in other contexts.2023-12-31T18:22:00ZThis version makes the title more explicit, revises the introduction and background, corrects some minor errors, and adds a few clarificationsMichael MolloyPawel PralatGregory B. Sorkinhttp://arxiv.org/abs/2401.00554v2Asymptotic Stability for Relativistic Vlasov-Maxwell-Landau System in Bounded Domain2024-12-15T23:52:19ZThe control of plasma-wall interactions is crucial to fusion devices from both physical and mathematical perspectives. It is well known that a magnetic field satisfying the classical perfect conducting conditions at the wall, $$
\mathbf{E} \times n_x = 0, \quad \mathbf{B} \cdot n_x = 0, $$ plays an important role in fusion plasma dynamics studies. Since the early 1990s, it has been understood that the Lorentz force can penetrate into the domain at the boundary and create a singularity. Consequently, the uniqueness for any nonlinear kinetic plasma models in the presence of a perfectly conducting boundary remained open until our recent local well-posedness result. In this paper, we finally establish a global well-posedness theory for the relativistic Vlasov-Maxwell-Landau system in a general $3D$ domain with a specularly reflective, perfectly conducting boundary.2023-12-31T17:54:07ZMajor revisions include the addition of a global well-posedness result. Further improvements are as follows: enhanced control of instant energies, more precise decay estimates of instant energies, and a significantly improved presentation. The manuscript now comprises 99 pagesHongjie DongYan GuoTimur Yastrzhembskiyhttp://arxiv.org/abs/2401.00573v1Proximal quantum control of spin and spin ensemble with highly localized control field from skyrmions2023-12-31T19:23:21ZSelective control of individual spin qubits is needed for scalable quantum computing based on spin states. Achieving high-fidelity in both single and two-qubit gates, essential components of universal quantum computers, necessitates highly localized control fields. These fields must be capable of addressing specific spin qubits while minimizing gate errors and cross-talk in adjacent qubits. Overcoming the challenge of generating a localized radio-frequency magnetic field, in the absence of elementary magnetic monopoles, we introduce a technique that combines divergent and convergent nanoscale magnetic skyrmions. This approach produces a precise control field that manipulates spin qubits with high fidelity. We propose the use of 2D skyrmions, which are 2D analogues of 3D hedgehog structures. The latter are emergent magnetic monopoles, but difficult to fabricate. The 2D skyrmions, on the other hand, can be fabricated using standard semiconductor foundry processes. Our comparative analysis of the density matrix evolution and gate fidelities in scenarios involving proximal skyrmions and nanomagnets indicates potential gate fidelities surpassing 99.95% for π/2-gates and 99.90% for π-gates. Notably, the skyrmion configuration generates a significantly lower field on neighboring spin qubits, i.e. 15 times smaller field on a neighboring qubit compared to nanomagnets that produces the same field at the controlled qubit, making it a more suitable candidate for scalable quantum control architectures by reducing disturbances in adjacent qubits.2023-12-31T19:23:21ZMd Fahim F ChowdhuryMohamad NiknamMd Mahadi RajibLouis S. BouchardJayasimha Atulasimha10.1103/PhysRevApplied.22.064077http://arxiv.org/abs/2401.00598v1Irreducible Maps and Isomorphisms of Boolean Algebras of Regular Open Sets and Regular Ideals2023-12-31T22:33:15ZLet $π: Y\rightarrow X$ be a continuous surjection between compact Hausdorff spaces $Y$ and $X$ which is irreducible in the sense that if $F\subsetneq Y$ is closed, then $π(F)\neq X$. We exhibit isomorphisms between various Boolean algebras associated to this data: the regular open sets of $X$, the regular open sets of $Y$, the regular ideals of $C(X)$ and the regular ideals of $C(Y)$.
We call $X$ and $Y$ Boolean equivalent if the regular open sets of $X$ and the regular open sets of $Y$ are isomorphic Boolean algebras. We give a characterization of when two compact metrizable spaces are Boolean equivalent; this characterization may be viewed as a topological version of the characterization of standard Borel spaces.2023-12-31T22:33:15Z13 pagesProc. Amer. Math. Soc. 153 (2025), 2713-2727David R. Pitts10.1090/proc/17234http://arxiv.org/abs/2401.00501v1Painting Taylor vortices with cellulose nanocrystals: supercritical spectral dynamics2023-12-31T13:47:40ZWe study the flow stability and spatio-temporal spectral dynamics of cellulose nanocrystal (CNC) suspensions in a custom Taylor-Couette flow cell using the intrinsic shear induced birefringence and liquid crystalline properties of CNC suspensions for flow visualizations for the first time. The analysis is performed at constant ramped speed inputs of the independently rotating cylinders for several cases ranging from only inner or outer rotating cylinders to three counter-rotation cases. All CNC suspensions have measurable elastic and shear thinning, both increasing with CNC concentration. We show that the flow patterns recorded are essentially Newtonian-like, with non-Newtonian effects ranging from a decrease in wavenumbers to altering the critical parameters for the onset of instability modes. Outer cylinder rotation flow cases are stable for all concentrations whereas inner cylinder rotation flow cases transition to axisymmetric and azimuthally periodic secondary flows. However, unstable counter-rotation cases become unstable to asymmetric spiral modes. With increasing CNC concentration a counter-rotation case was found where azimuthally periodic wavy patterns transition to asymmetric spiral modes. In contrast to polymeric solutions of similar low to moderate elasticity and shear thinning, the shear-thinning region of CNC suspensions is expected to lead to the breakdown of the chiral nematic phase, whose elastic constants constitute the dominant structural elasticity mechanism. Thus, we interpret the Taylor-Couette stability of the CNC suspensions as dominated by their shear-thinning character due to the expected loss of elasticity in nonlinear flow conditions.2023-12-31T13:47:40ZReza GhanbariSajjad PashazadehKesavan SekarKim NygårdAnn TerryMarianne LiebiAleksandar MaticRoland Kádár10.1063/5.0195130http://arxiv.org/abs/2401.00506v2Andreev bound states in Josephson junctions of semi-Dirac semimetals2024-04-08T19:19:55ZWe consider a Josephson junction built with the two-dimensional semi-Dirac semimetal, which features a hybrid of linear and quadratic dispersion around a nodal point. We model the weak link between the two superconducting regions by a Dirac delta potential because it mimics the thin-barrier-limit of a superconductor-barrier-superconductor configuration. Assuming a homogeneous pairing in each region, we set up the BdG formalism for electronlike and holelike quasiparticles propagating along the quadratic-in-momentum dispersion direction. This allows us to compute the discrete bound-state energy spectrum $\varepsilon $ of the subgap Andreev states localized at the junction. In contrast with the Josephson effect investigated for propagation along linearly dispersing directions, we find a pair of doubly degenerate Andreev bound states. Using the dependence of $\varepsilon $ on the superconducting phase difference $φ$, we compute the variation of Josephson current as a function of $φ$.2023-12-31T13:57:09Zjournal version; companion paper of arXiv:2312.16164Physica B: Condensed Matter 683, 415918 (2024)Ipsita Mandal10.1016/j.physb.2024.415918http://arxiv.org/abs/2401.00532v1On the Necessity of Metalearning: Learning Suitable Parameterizations for Learning Processes2023-12-31T16:24:03ZIn this paper we will discuss metalearning and how we can go beyond the current classical learning paradigm. We will first address the importance of inductive biases in the learning process and what is at stake: the quantities of data necessary to learn. We will subsequently see the importance of choosing suitable parameterizations to end up with well-defined learning processes. Especially since in the context of real-world applications, we face numerous biases due, e.g., to the specificities of sensors, the heterogeneity of data sources, the multiplicity of points of view, etc. This will lead us to the idea of exploiting the structuring of the concepts to be learned in order to organize the learning process that we published previously. We conclude by discussing the perspectives around parameter-tying schemes and the emergence of universal aspects in the models thus learned.2023-12-31T16:24:03ZMassinissa HamidiAomar Osmanihttp://arxiv.org/abs/2401.00547v2On Learning for Ambiguous Chance Constrained Problems2024-02-11T06:07:17ZWe study chance constrained optimization problems $\min_x f(x)$ s.t. $P(\left\{ θ: g(x,θ)\le 0 \right\})\ge 1-ε$ where $ε\in (0,1)$ is the violation probability, when the distribution $P$ is not known to the decision maker (DM). When the DM has access to a set of distributions $\mathcal{U}$ such that $P$ is contained in $\mathcal{U}$, then the problem is known as the ambiguous chance-constrained problem \cite{erdougan2006ambiguous}. We study ambiguous chance-constrained problem for the case when $\mathcal{U}$ is of the form $\left\{μ:\frac{μ(y)}{ν(y)}\leq C, \forall y\inΘ, μ(y)\ge 0\right\}$, where $ν$ is a ``reference distribution.'' We show that in this case the original problem can be ``well-approximated'' by a sampled problem in which $N$ i.i.d. samples of $θ$ are drawn from $ν$, and the original constraint is replaced with $g(x,θ_i)\le 0,~i=1,2,\ldots,N$. We also derive the sample complexity associated with this approximation, i.e., for $ε,δ>0$ the number of samples which must be drawn from $ν$ so that with a probability greater than $1-δ$ (over the randomness of $ν$), the solution obtained by solving the sampled program yields an $ε$-feasible solution for the original chance constrained problem.2023-12-31T17:25:43ZWe have "not considered the uniform bound" for violation probabilities corresponding to the set of distributions in the ambiguity setA Ch MadhusudanaraoRahul Singhhttp://arxiv.org/abs/2401.00551v1A Generalist FaceX via Learning Unified Facial Representation2023-12-31T17:41:48ZThis work presents FaceX framework, a novel facial generalist model capable of handling diverse facial tasks simultaneously. To achieve this goal, we initially formulate a unified facial representation for a broad spectrum of facial editing tasks, which macroscopically decomposes a face into fundamental identity, intra-personal variation, and environmental factors. Based on this, we introduce Facial Omni-Representation Decomposing (FORD) for seamless manipulation of various facial components, microscopically decomposing the core aspects of most facial editing tasks. Furthermore, by leveraging the prior of a pretrained StableDiffusion (SD) to enhance generation quality and accelerate training, we design Facial Omni-Representation Steering (FORS) to first assemble unified facial representations and then effectively steer the SD-aware generation process by the efficient Facial Representation Controller (FRC). %Without any additional features, Our versatile FaceX achieves competitive performance compared to elaborate task-specific models on popular facial editing tasks. Full codes and models will be available at https://github.com/diffusion-facex/FaceX.2023-12-31T17:41:48ZProject page: https://diffusion-facex.github.io/Yue HanJiangning ZhangJunwei ZhuXiangtai LiYanhao GeWei LiChengjie WangYong LiuXiaoming LiuYing Taihttp://arxiv.org/abs/2401.00567v1Mean ergodic theorems in $L^r(μ)$ and $H^r(\mathbb T)$, $0<r<1$2023-12-31T19:05:09ZLet $T$ be the Koopman operator of a measure preserving transformation $θ$ of a probability space $(X,Σ,μ)$. We study the convergence properties of the averages $M_nf:=\frac1n\sum_{k=0}^{n-1}T^kf$ when $f \in L^r(μ)$, $0<r<1$. We prove that if $\int |M_nf|^r dμ\to 0$, then $f \in \overline{(I-T)L^r}$, and show that the converse fails whenever $θ$ is ergodic aperiodic. When $θ$ is invertible ergodic aperiodic, we show that for $0<r<1$ there exists $f_r \in (I-T)L^r$ for which $M_nf_r$ does not converge a.e. (although $\int |M_nf|^r dμ\to 0$). We further establish that for $1 \leq p <\frac{1}{r},$ there is a dense $G_δ$ subset ${\mathcal F}\subset L^p(X,μ)$ such that $\limsup_n \frac{|T^nh|}{n^r}=\infty$ a.e. for any $h \in {\mathcal F}$.2023-12-31T19:05:09Z18 pages, 25 references, 3 Lemmas, 6 Theorems, 12 Propositions and 7 Remarksel Houcein el AbdalaouiMichael Linhttp://arxiv.org/abs/2401.00581v1Measurement and analysis of the Doppler broadened energy spectra of annihilation gamma radiation originating from clean and adsorbate-covered surfaces2023-12-31T20:06:06ZWe present measurements and theoretical modeling demonstrating the capability of Doppler Broadened annihilation gamma Spectroscopy (DBS) to provide element-specific information from the topmost atomic layer of surfaces that are either clean or covered with adsorbates or thin films. Our measurements show that the energy spectra of Doppler-shifted annihilation gamma photons emitted following the annihilation of positrons from the topmost atomic layers of clean gold (Au) and copper (Cu) differ significantly. With the aid of the positron annihilation-induced Auger electron spectroscopy (PAES) performed simultaneously with DBS, we show that measurable differences between the Doppler broadened gamma spectra from Au and Cu surfaces in the high energy region of the gamma spectra can be used for the quantification of surface chemical composition. Modeling the measured Doppler spectra from clean Au and Cu surfaces using gamma spectra obtained from ab initio calculations after considering the detector energy resolution and surface positronium formation pointed to an increase in the relative contribution of gamma from positron annihilation with valence shell electrons. The fit result also suggests that the surface-trapped positrons predominantly annihilated with the delocalized valence shell (s and p) electrons that extended into the vacuum as compared to the highly localized d electrons. Simultaneous DBS and PAES measurements from adsorbate (sulfur, oxygen, carbon) or thin film (selenium (Se), graphene) covered Cu surface showed that it is possible to distinguish and quantify the surface adsorbate and thin-film composition just based on DBS. DBS of elemental surfaces presents a promising avenue for developing a characterization tool that can be used to probe external and internal surfaces that are inaccessible by conventional surface science techniques.2023-12-31T20:06:06ZS. LotfimaranglooV. A. ChirayathP. A. SterneH. MahdyR. W. GladenJ. DriscollM. RooksM. ChryslerA. R. KoymenJ. AsaadiA. H. Weisshttp://arxiv.org/abs/2401.00565v3Photometric Objects Around Cosmic Webs (PAC). VI. High Satellite Fraction of Quasars2024-05-15T06:53:14ZThe Photometric objects Around Cosmic webs (PAC) approach developed in Xu et al. (2022b) has the advantage of making full use of spectroscopic and deeper photometric surveys. With the merits of PAC, the excess surface density $\bar{n}_2w_{\rm{p}}$ of neighboring galaxies can be measured down to stellar mass $10^{10.80}\,M_{\odot}$ around quasars at redshift $0.8<z_{\rm{s}}<1.0$, with the data from the Sloan Digital Sky Survey IV (SDSS-IV) extended Baryon Oscillation Spectroscopic Survey (eBOSS) and the Dark Energy Spectroscopic Instrument (DESI) Legacy Imaging Surveys. We find that $\bar{n}_2w_{\rm{p}}$ generally increases quite steeply with the decrease of the separation. Using subhalo abundance matching method, we can accurately model the $\bar{n}_2w_{\rm{p}}$ both on small and large scales. We show that the steep increase of the $\bar{n}_2w_{\rm{p}}$ towards the quasars requires that a large fraction $f_{\mathrm{sate}}=0.29_{-0.06}^{+0.05}$ of quasars should be satellites in massive halos, and find that this fraction measurement is insensitive to the assumptions of our modeling. This high satellite fraction indicates that the subhalos have nearly the same probability to host quasars as the halos for the same (infall) halo mass, and the large scale environment has negligible effect on the quasar activity. We show that even with this high satellite fraction, each massive halo on average does not host more than one satellite quasar due to the sparsity of quasars.2023-12-31T19:00:04Z15 pages, 11 figures, 2 tables, accepted for publication in the Astrophysical JournalThe Astrophysical Journal, 967:17 (13pp), 2024 May 20Shanquan GuiSJTUKun XuSJTU, DurhamY. P. JingSJTU, TDLIDonghai ZhaoSHAO, SJTUHongyu GaoSJTU10.3847/1538-4357/ad3b96http://arxiv.org/abs/2401.00595v3State of What Art? A Call for Multi-Prompt LLM Evaluation2024-05-06T10:20:26ZRecent advances in large language models (LLMs) have led to the development of various evaluation benchmarks. These benchmarks typically rely on a single instruction template for evaluating all LLMs on a specific task. In this paper, we comprehensively analyze the brittleness of results obtained via single-prompt evaluations across 6.5M instances, involving 20 different LLMs and 39 tasks from 3 benchmarks. To improve robustness of the analysis, we propose to evaluate LLMs with a set of diverse prompts instead. We discuss tailored evaluation metrics for specific use cases (e.g., LLM developers vs. developers interested in a specific downstream task), ensuring a more reliable and meaningful assessment of LLM capabilities. We then implement these criteria and conduct evaluations of multiple models, providing insights into the true strengths and limitations of current LLMs.2023-12-31T22:21:36ZAccepted at TACL; pre-MIT Press publication versionMoran MizrahiGuy KaplanDan MalkinRotem DrorDafna ShahafGabriel Stanovskyhttp://arxiv.org/abs/2401.00597v2Local dual spaces and primary decomposition2024-11-30T15:21:40ZGeneralizing the concept of the Macaulay inverse system, we introduce a way to describe localizations of an ideal in a polynomial ring. This leads to an approach to the differential primary decomposition as a description of the affine scheme defined by the ideal.2023-12-31T22:28:47Z7 pagesJournal of Symbolic Computation (2024): 102400Justin ChenMarc HärkönenAnton Leykinhttp://arxiv.org/abs/2401.00560v2Monopole acceleration in intergalactic magnetic fields2024-10-24T15:32:50ZWe provide a comprehensive analysis of the acceleration of magnetic monopoles in intergalactic magnetic fields. We demonstrate that monopoles with intermediate to low masses can be accelerated to relativistic velocities. This can significantly affect direct and indirect searches for magnetic monopoles. As an example, we show that the Parker bound is relaxed in the presence of intergalactic fields. We also find that a cosmic population of monopoles can produce significant backreaction on the intergalactic fields.2023-12-31T18:22:10Z23 pages, 5 figuresPhysics of the Dark Universe, 2024, 101704, ISSN 2212-6864Daniele PerriKyrilo BondarenkoMichele DoroTakeshi Kobayashi10.1016/j.dark.2024.101704http://arxiv.org/abs/2401.00511v2Binary AdS black holes coupled to a bath in Type IIB2024-08-31T08:05:08ZWe construct Type IIB string theory setups which, via double holography, realize two gravitational systems in separate AdS spaces which interact with each other and with a non-gravitational bath. We employ top-down string theory solutions with concrete field theory duals in the form of 4d $\mathcal N=4$ SYM BCFTs and a first-principles notion of double holography. The setups are used to realize pairs of `near' and `far' black holes from the perspective of the bath, which exchange Hawking radiation with each other and radiate into the bath. We identify three phases for the entropy in the bath characterized as no island, partial island and full island, and discuss the entropy curves. The setups differ from the black hole binaries observed in gravitational wave experiments but may capture certain aspects.2023-12-31T14:35:58Z28 pages, 10 figures; v2: references added, published versionEvan DeddoLeopoldo A. Pando ZayasChristoph F. Uhlemann10.1007/JHEP05(2024)120http://arxiv.org/abs/2401.00594v3Fast and Scalable Beamforming for RIS-Assisted Downlink Multi-group Multicasting2025-07-04T15:07:32ZThis paper considers downlink multi-group multicasting via beamforming facilitated by a reconfigurable intelligent surface (RIS). We develop a fast and scalable algorithm for the joint base station (BS) and RIS beamforming optimization to minimize the transmit power while meeting user quality-of-service (QoS) targets. By analyzing the structure of the QoS constraints, we reformulate the problem and show that the joint beamforming optimization inherently consists of a multicast beamforming QoS problem for the BS and a passive multicast beamforming max-min-fair (MMF) problem for the RIS. We propose a fast alternating multicast beamforming (AMBF) algorithm to effectively solve the two subproblems alternatingly. For the BS multicast subproblem, we utilize the optimal multicast beamforming structure to efficiently determine the BS beamformers. For the RIS multicast subproblem, we reformulate the MMF problem and apply a first-order projected subgradient algorithm (PSA), which yields simple closed-form updates. The computational complexity of the AMBF algorithm grows linearly with the number of RIS elements and BS antennas. We further consider joint BS and RIS beamforming for the weighted MMF design objective, subject to the BS transmit power budget. We propose an alternating PSA (APSA) fast algorithm to compute the beamforming solutions for the BS and RIS. APSA consists of only closed-form updates per iteration, yielding linear computational complexity in the number of RIS elements and BS antennas. Simulation results show the efficacy of our proposed algorithms in terms of performance and computational cost compared to alternative methods.2023-12-31T22:17:38Z14 pages, 8 figuresMohammad EbrahimiMin DongMitra Hekmathttp://arxiv.org/abs/2401.00529v3GraphGPT: Generative Pre-trained Graph Eulerian Transformer2025-06-06T09:46:30ZWe introduceGraphGPT, a novel self-supervised generative pre-trained model for graph learning based on the Graph Eulerian Transformer (GET). First, we propose GET, which combines a standard transformer encoder or decoder architecture with an innovative graph-to-sequence transformation method. This method converts graphs or sampled subgraphs into sequences of tokens representing nodes, edges, and attributes in a reversible manner using Eulerian paths. We pre-train GET using either of the two self-supervised tasks: next-token prediction (NTP) and scheduled masked-token prediction (SMTP). The pre-trained model is then fine-tuned for downstream tasks such as graph-, edge-, and node-level prediction. Despite its simplicity, GraphGPT achieves performance comparable to or surpassing state-of-the-art methods on multiple large-scale Open Graph Benchmark (OGB) datasets. It demonstrates exceptional results on the molecular property prediction dataset PCQM4Mv2 and the protein-protein interaction dataset ogbl-ppa. Notably, generative pre-training enables scaling GraphGPT to 2 billion parameters while maintaining performance gains - a breakthrough that overcomes the scalability limitations of traditional Graph Neural Networks (GNNs) and prior graph transformers (GTs). To advance research in graph foundation models and facilitate scientific discovery in chemistry, materials science, and related fields, we will release the source code (https://github.com/alibaba/graph-gpt) and pre-trained checkpoints.2023-12-31T16:19:30Z9 pagesICML2025Qifang ZhaoWeidong RenTianyu LiHong LiuXingsheng HeXiaoxiao Xuhttp://arxiv.org/abs/2401.00541v1Ideals and their Fitting ideals2023-12-31T17:07:20ZFor an ideal $I$ in a Noetherian ring $R$, the Fitting ideals $\textrm{Fitt}_j(I)$ are studied. We discuss the question of when $\textrm{Fitt}_j(I)=I$ or $\sqrt{\textrm{Fitt}_j(I)}=\sqrt{I}$ for some $j$. A classical case is the Hilbert-Burch theorem when $j=1$ and $I$ is a perfect ideal of grade $2$ in a local ring.2023-12-31T17:07:20ZMath. Proc. Camb. Phil. Soc. 179 (2025) 609-621David EisenbudAntonino FicarraJürgen HerzogSomayeh Moradi10.1017/S0305004125101394http://arxiv.org/abs/2401.00592v5Majority voting is not good for heaven or hell, with mirrored performance2025-12-20T19:03:51ZWithin the ViSE (Voting in Stochastic Environment) model, we study the effectiveness of majority voting in various environments. As shown by the pit-of-losses paradox identified in previous work, majority decisions in apparently hostile environments tend to reduce the capital of society. In such cases, the simple social decision rule of ``rejecting all proposals without voting'' outperforms majority voting. In this paper, we identify another pit of losses appearing in favorable environments; here, the simple social decision rule of ``accepting all proposals without voting'' is superior to majority voting. We prove that, under a version of simple majority called symmetrized majority and under the antisymmetry of the voting body, this second pit of losses is a mirror image of the one arising in hostile environments, and we explain this phenomenon. Technically, we consider a voting society consisting of individualists who support all proposals that increase their personal capital and a group (or groups) whose members vote to increase their group's wealth. According to the key lemma, the expected capital gain of each agent under the social decision rule when the random gain generator is $X$ with mean $μ>0$ exceeds their expected gain under the reflected generator $-X$ by exactly $μ$. This extends to location-scale families of generators with distributions symmetric about their mean. This result reveals a mirror symmetry in the performance of the symmetrized majority rule relative to a baseline rule. The baseline rule accepts all proposals in favorable environments and rejects them in unfavorable (hostile) ones.2023-12-31T21:59:40Z19 pages, 4 figures. Submitted to a journal. Compared to the previous version, the results have been generalizedPavel ChebotarevVadim Afonkinhttp://arxiv.org/abs/2401.00527v2Sub-Poissonian estimates for exponential moments of additive functionals over pairs of particles with respect to determinantal and symplectic Pfaffian point processes governed by entire functions2025-12-29T15:11:08ZThe aim of this note is to estimate the tail of the distribution of the number of particles in an interval under determinantal and Pfaffian point processes. The main result of the note is that the square of the number of particles under the determinantal point process whose correlation kernel is an entire function of finite order has sub-Poissonian tails. The same result also holds in the symplectic Pfaffian case. As a corollary, sub-Poissonian estimates are also obtained for exponential moments of additive functionals over pairs of particles.2023-12-31T16:10:08Z18 pages; references have been updatedMoscow Mathematical Journal, 23:4(2023), 463-478Alexander I. Bufetovhttp://arxiv.org/abs/2401.00513v2Non-existence of three non-coalescing infinite geodesics with the same direction in the directed landscape2024-01-30T15:14:48ZIt is believed that for metric-like models in the KPZ class the following property holds: with probability one, starting from any point, there are at most two semi-infinite geodesics with the same direction that do not coalesce. Until now, such a result was only proved for one model - exponential LPP (Coupier 11') using its inherent connection to the totally asymmetric exclusion process. We prove that the above property holds for the directed landscape, the universal scaling limit of models in the KPZ class. Our proof reduces the problem to one on line ensembles and therefore paves the way to show similar results for other metric-like models in the KPZ class. Finally, combining our result with the ones in (Busani, Seppalainen,Sorensen 22', Bhatia 23') we obtain the full qualitative geometric description of infinite geodesics in the directed landscape.2023-12-31T14:43:21Z38 pages. Section 7 has shorten considerablyOfer Busanihttp://arxiv.org/abs/2401.00515v2Empirical Analysis of Vulnerabilities Life Cycle in Golang Ecosystem2024-01-17T06:16:52ZOpen-source software (OSS) greatly facilitates program development for developers. However, the high number of vulnerabilities in open-source software is a major concern, including in Golang, a relatively new programming language. In contrast to other commonly used OSS package managers, Golang presents a distinctive feature whereby commits are prevalently used as dependency versions prior to their integration into official releases. This attribute can prove advantageous to users, as patch commits can be implemented in a timely manner before the releases. However, Golang employs a decentralized mechanism for managing dependencies, whereby dependencies are upheld and distributed in separate repositories. This approach can result in delays in the dissemination of patches and unresolved vulnerabilities.
To tackle the aforementioned concern, a comprehensive investigation was undertaken to examine the life cycle of vulnerability in Golang, commencing from its introduction and culminating with its rectification. To this end, a framework was established by gathering data from diverse sources and systematically amalgamating them with an algorithm to compute the lags in vulnerability patching. It turned out that 66.10% of modules in the Golang ecosystem were affected by vulnerabilities. Within the vulnerability life cycle, we found two kinds of lag impeding the propagation of vulnerability fixing. By analyzing reasons behind non-lagged and lagged vulnerabilities, timely releasing and indexing patch versions could significantly enhance ecosystem security.2023-12-31T14:53:51Z12 pages, 10 figures, 2024 IEEE/ACM 46th International Conference on Software Engineering (ICSE '24)Jinchang HuCollege of Command and Control Engineering, Army Engineering University of PLA, NanJing, ChinaLyuye ZhangContinental-NTU Corporate Lab, Nanyang Technological University, Singapore, SingaporeChengwei LiuContinental-NTU Corporate Lab, Nanyang Technological University, Singapore, SingaporeSen YangAcademy of Military Science, BeiJing, ChinaSong HuangCollege of Command and Control Engineering, Army Engineering University of PLA, NanJing, ChinaYang LiuContinental-NTU Corporate Lab, Nanyang Technological University, Singapore, Singapore10.1145/3597503.3639230http://arxiv.org/abs/2401.00517v1Detecting Imprinting and Maternal Effects Using Monte Carlo Expectation Maximization Algorithm2023-12-31T15:05:46ZNumerous statistical methods have been developed to explore genomic imprinting and maternal effects, which are causes of parent-of-origin patterns in complex human diseases. However, most of them either only model one of these two confounded epigenetic effects, or make strong yet unrealistic assumptions about the population to avoid over-parameterization. A recent partial likelihood method (LIME) can identify both epigenetic effects based on case-control family data without those assumptions. Theoretical and empirical studies have shown its validity and robustness. However, because LIME obtains parameter estimation by maximizing partial likelihood, it is interesting to compare its efficiency with full likelihood maximizer. To overcome the difficulty in over-parameterization when using full likelihood, in this study we propose a Monte Carlo Expectation Maximization (MCEM) method to detect imprinting and maternal effects jointly. Those unknown mating type probabilities, the nuisance parameters, can be considered as latent variables in EM algorithm. Monte Carlo samples are used to numerically approximate the expectation function that cannot be solved algebraically. Our simulation results show that though this MCEM algorithm takes longer computational time, and can give higher bias in some simulations compared to LIME, it can generally detect both epigenetic effects with higher power and smaller standard error which demonstrates that it can be a good complement of LIME method.2023-12-31T15:05:46ZPooya AavaniAlexandre TrindadeFangyuan Zhanghttp://arxiv.org/abs/2401.00528v1On the Stiefel-Whitney classes of GKM manifolds2023-12-31T16:17:55ZWe show that under standard assumptions on the isotropy groups of an integer GKM manifold, the equivariant Stiefel-Whitney classes of the action are determined by the GKM graph. This is achieved via a GKM-style description of the equivariant cohomology with coefficients in a finite field $\mathbb Z_{p}$ even though in this setting the restriction map to the fixed point set is not necessarily injective. This closes a gap in our argument why the GKM graph of a $6$-dimensional integer GKM manifold determines its nonequivariant diffeomorphism type. We introduce combinatorial Stiefel-Whitney classes of GKM graphs and use them to derive a nontrivial obstruction to realizability of GKM graphs in dimension $8$ and higher.2023-12-31T16:17:55Z19 pages, comments are welcome!Oliver GoertschesPanagiotis KonstantisLeopold Zollerhttp://arxiv.org/abs/2401.00534v1Financial Time-Series Forecasting: Towards Synergizing Performance And Interpretability Within a Hybrid Machine Learning Approach2023-12-31T16:38:32ZIn the realm of cryptocurrency, the prediction of Bitcoin prices has garnered substantial attention due to its potential impact on financial markets and investment strategies. This paper propose a comparative study on hybrid machine learning algorithms and leverage on enhancing model interpretability. Specifically, linear regression(OLS, LASSO), long-short term memory(LSTM), decision tree regressors are introduced. Through the grounded experiments, we observe linear regressor achieves the best performance among candidate models. For the interpretability, we carry out a systematic overview on the preprocessing techniques of time-series statistics, including decomposition, auto-correlational function, exponential triple forecasting, which aim to excavate latent relations and complex patterns appeared in the financial time-series forecasting. We believe this work may derive more attention and inspire more researches in the realm of time-series analysis and its realistic applications.2023-12-31T16:38:32ZShun LiuKexin WuChufeng JiangBin HuangDanqing Mahttp://arxiv.org/abs/2401.00545v1Production of s-process elements in AGB stars as revealed by Gaia GSP-spec abundances2023-12-31T17:16:47ZThe recent parameterisation by the GSP-spec module of Gaia/RVS spectra has produced an homogeneous catalogue of about 174,000 AGB stars. Among the 13 chemical elements presented in this catalogue, the abundance of 2 of them (Ce and Nd) have been estimated in most of these AGBs. These 2 species formed by slow n-captures in the interior of low- and intermediate-mass stars, belong to the family of 2nd-peak s-process elements. We defined a working sample of 19,544 AGB stars with high-quality Ce and/or Nd abundances, selected by applying a specific combination of the GSP-spec quality flags. We compared these abundances with the yield production predicted by AGB stars evolutionary models. We found a good correlation between the Ce and Nd abundances, confirming the high quality of the derived abundances and that these species indeed belong to the same s-process family. We also found higher Ce and Nd abundances for more evolved AGB stars of similar metallicity, illustrating the successive mixing episodes enriching the AGB star surface. We then compared the observed Ce and Nd abundances with the FRUITY and Monash AGB yields and found that the higher Ce and Nd abundances cannot be explained by AGB stars of masses higher than 5Msun. In contrast, the yields predicted by both models for AGB stars with an initial mass between ~1.5 and ~2.5Mssun and metallicities between ~-0.5 and ~0.0dex are fully compatible with the observed GSP-spec abundances. This work based on the largest catalogue of high-quality second-peak s-element abundances in O-rich AGB stars allows evolutionary models to be constrained and confirms the fundamental role played by low- and intermediate-mass stars in the enrichment of the Universe in these chemical species.2023-12-31T17:16:47ZAccepted by A&A in october, 2023G. ContursiP. de LavernyA. Recio-BlancoP. A. PalicioC. Abiahttp://arxiv.org/abs/2401.00556v1The evaluation of a definite integral by the method of brackets illustrating its flexibility2023-12-31T18:04:35ZThe method of brackets is an procedure to evaluate definite integrals. It is based on a small number of operational rules. The flexibility of this method is illustrated with the evaluation of an integral involving the Bessel K0 function and the exponential integral. Several proofs are presented.2023-12-31T18:04:35ZIvan GonzalezJohn Lopez SantanderVictor H. Mollhttp://arxiv.org/abs/2401.00557v1Sobolev spaces on hypergroups Gelfand pairs2023-12-31T18:08:18ZThis paper introduces Sobolev spaces over Gelfand pairs in the framework of hypergroups. The Sobolev spaces in question are constructed from the Fourier transform on hypergroup Gelfand pairs. Mainly, the paper focuses on the investigation of Sobolev embedding results.2023-12-31T18:08:18ZKy T. BatakaMurphy E. EgweYaogan Mensahhttp://arxiv.org/abs/2401.00566v1Change point analysis -- the empirical Hankel transform approach2023-12-31T19:03:01ZIn this study, we introduce the first-of-its-kind class of tests for detecting change points in the distribution of a sequence of independent matrix-valued random variables. The tests are constructed using the weighted square integral difference of the empirical orthogonal Hankel transforms. The test statistics have a convenient closed-form expression, making them easy to implement in practice. We present their limiting properties and demonstrate their quality through an extensive simulation study. We utilize these tests for change point detection in cryptocurrency markets to showcase their practical use. The detection of change points in this context can have various applications in constructing and analyzing novel trading systems.2023-12-31T19:03:01Z13 pagesŽikica LukićBojana Miloševićhttp://arxiv.org/abs/2401.00575v1Neural Networks Against (and For) Self-Training: Classification with Small Labeled and Large Unlabeled Sets2023-12-31T19:25:34ZWe propose a semi-supervised text classifier based on self-training using one positive and one negative property of neural networks. One of the weaknesses of self-training is the semantic drift problem, where noisy pseudo-labels accumulate over iterations and consequently the error rate soars. In order to tackle this challenge, we reshape the role of pseudo-labels and create a hierarchical order of information. In addition, a crucial step in self-training is to use the classifier confidence prediction to select the best candidate pseudo-labels. This step cannot be efficiently done by neural networks, because it is known that their output is poorly calibrated. To overcome this challenge, we propose a hybrid metric to replace the plain confidence measurement. Our metric takes into account the prediction uncertainty via a subsampling technique. We evaluate our model in a set of five standard benchmarks, and show that it significantly outperforms a set of ten diverse baseline models. Furthermore, we show that the improvement achieved by our model is additive to language model pretraining, which is a widely used technique for using unlabeled documents. Our code is available at https://github.com/p-karisani/RST.2023-12-31T19:25:34ZACL Findings 2023Payam Karisanihttp://arxiv.org/abs/2401.00576v1High-Order, Implicit Time Integration of Discrete, Chaotic Dynamical Systems2023-12-31T19:35:34ZA wide range of implicit time integration methods, including multi-step, implicit Runge-Kutta, and Galerkin finite-time element schemes, is evaluated in the context of chaotic dynamical systems. The schemes are applied to solve the Lorenz equations, the equation of motion of a Duffing oscillator, and the Kuramoto-Sivashinsky system, with the goal of finding the most computationally efficient method that results in the least expensive model for a chosen level of accuracy. It is found that the quasi-period of a chaotic system strongly limits the time-step size that can be used in the simulations, and all schemes fail once the time-step size reaches a significant fraction of that period. In these conditions, the computational cost per time-step becomes one of the most important factors determining the efficiency of the schemes. The cheaper, second-order schemes are shown to have an advantage over the higher-order schemes at large time-step sizes, with one possible exception being the fourth-order continuous Galerkin scheme. The higher-order schemes become more efficient than the lower-order schemes as accuracy requirements tighten. If going beyond the second-order is necessary for reasons other than computational efficiency, the fourth-order methods are shown to perform better than the third-order ones at all time-step sizes.2023-12-31T19:35:34Z38 pages, 15 figuresViktoriya MorozovaJames G. CoderKevin Holsthttp://arxiv.org/abs/2401.00577v1Bayesian Inference for Contemporary Lattice Quantum Field Theory2023-12-31T19:56:42ZBayesian inference provides a rigorous framework to encapsulate our knowledge and uncertainty regarding various physical quantities in a well-defined and self-contained manner. Utilising modern tools, such Bayesian models can be constructed with a remarkable flexibility, leaving us totally free to carefully choose which assumption should be strictly enforced and which should on the contrary be relaxed. The practical evaluation of these assumptions, together with the data-driven selection or averaging of models, also appears in a very natural way. In this presentation, I discuss its application in the context of lattice QCD and its common statistical problems. As a concrete illustration, I present a few parametric and non-parametric hierarchical models applied to actual correlator data, from single exponential fits to spectral functions.2023-12-31T19:56:42Z10 pages, 6 figures, 40th International Symposium on Lattice Field TheoryJulien Frisonhttp://arxiv.org/abs/2401.00578v1Exact Error in Matrix Completion: Approximately Low-Rank Structures and Missing Blocks2023-12-31T19:58:08ZWe study the completion of approximately low rank matrices with entries missing not at random (MNAR). In the context of typical large-dimensional statistical settings, we establish a framework for the performance analysis of the nuclear norm minimization ($\ell_1^*$) algorithm. Our framework produces \emph{exact} estimates of the worst-case residual root mean squared error and the associated phase transitions (PT), with both exhibiting remarkably simple characterizations. Our results enable to {\it precisely} quantify the impact of key system parameters, including data heterogeneity, size of the missing block, and deviation from ideal low rankness, on the accuracy of $\ell_1^*$-based matrix completion. To validate our theoretical worst-case RMSE estimates, we conduct numerical simulations, demonstrating close agreement with their numerical counterparts.2023-12-31T19:58:08Z3 figures. arXiv admin note: text overlap with arXiv:2301.00793Agostino CapponiMihailo Stojnichttp://arxiv.org/abs/2401.00583v1Improving the Privacy and Practicality of Objective Perturbation for Differentially Private Linear Learners2023-12-31T20:32:30ZIn the arena of privacy-preserving machine learning, differentially private stochastic gradient descent (DP-SGD) has outstripped the objective perturbation mechanism in popularity and interest. Though unrivaled in versatility, DP-SGD requires a non-trivial privacy overhead (for privately tuning the model's hyperparameters) and a computational complexity which might be extravagant for simple models such as linear and logistic regression. This paper revamps the objective perturbation mechanism with tighter privacy analyses and new computational tools that boost it to perform competitively with DP-SGD on unconstrained convex generalized linear problems.2023-12-31T20:32:30ZRachel RedbergAntti KoskelaYu-Xiang Wanghttp://arxiv.org/abs/2401.00584v1Representing maps for semibounded forms and their Lebesgue type decompositions2023-12-31T20:35:55ZFor a semibounded sesquilinear form ${\mathfrak t}$ in a Hilbert space ${\mathfrak H}$ there exists a representing map $Q$ from ${\mathfrak H}$ to another Hilbert space ${\mathfrak K}$, such that ${\mathfrak t}[\varphi, ψ]-c(\varphi, ψ)=(Q\varphi,Qψ)$, $\varphi,ψ\in {\rm dom\,}{\mathfrak t}$, with $c \in {\mathbb R}$ a lower bound of ${\mathfrak t}$. Representing maps offer a simplifying tool to study general semibounded forms. By means of representing maps closedness, closability, and singularity of ${\mathfrak t}$ are immediately translated into the corresponding properties of the operator $Q$, and vice versa. Also properties of sum decompositions ${\mathfrak t}={\mathfrak t}_1+{\mathfrak t}_2$ of a nonnegative form ${\mathfrak t}$ with two other nonnegative forms ${\mathfrak t}_1$ and ${\mathfrak t}_2$ in ${\mathfrak H}$ can be analyzed by means of associated nonnegative contractions $K\in {\mathbf B}({\mathfrak K})$. This helps, for instance, to establish an explicit operator theoretic characterization for the summands ${\mathfrak t}_1$ and ${\mathfrak t}_2$ to be, or not to be, mutually singular. Such sum decompositions are used to study characteristic properties of the so-called Lebesgue type decompositions of semibounded forms ${\mathfrak t}$, where ${\mathfrak t}_1$ is closable and ${\mathfrak t}_2$ singular; in particular, this includes the Lebesgue decomposition of a semibounded form due to B. Simon. Furthermore, for a semibounded form ${\mathfrak t}$ with its representing map $Q$ it will be shown that the corresponding semibounded selfadjoint relation $Q^*Q^{**} +c$ is uniquely determined by a limit version of the classical representation theorem for the form ${\mathfrak t}$, being studied by W. Arendt and T. ter Elst in a sectorial context. Via representing maps a full treatment is given of the convergence of monotone sequences of semibounded forms.2023-12-31T20:35:55Z29 pagesSeppo HassiHenk de Snoohttp://arxiv.org/abs/2401.00599v1Sub-sampling of NMR Correlation and Exchange Experiments2023-12-31T22:33:42ZSub-sampling is applied to simulated $T_1$-$D$ NMR signals and its influence on inversion performance is evaluated. For this different levels of sub-sampling were employed ranging from the fully sampled signal down to only less than two percent of the original data points. This was combined with multiple sample schemes including fully random sampling, truncation and a combination of both. To compare the performance of different inversion algorithms, the so-generated sub-sampled signals were inverted using Tikhonov regularization, modified total generalized variation (MTGV) regularization, deep learning and a combination of deep learning and Tikhonov regularization. Further, the influence of the chosen cost function on the relative inversion performance was investigated. Overall, it could be shown that for a vast majority of instances, deep learning clearly outperforms regularization based inversion methods, if the signal is fully or close to fully sampled. However, in the case of significantly sub-sampled signals regularization yields better inversion performance than its deep learning counterpart with MTGV clearly prevailing over Tikhonov. Additionally, fully random sampling could be identified as the best overall sampling scheme independent of the inversion method. Finally, it could also be shown that the choice of cost function does vastly influence the relative rankings of the tested inversion algorithms highlighting the importance of choosing the cost function accordingly to experimental intentions.2023-12-31T22:33:42ZJulian B. B. BeckmannMick D. MantleAndrew J. SedermanLynn F. Gladdenhttp://arxiv.org/abs/2401.00591v1Magnetic dipole $γ$-ray strength functions of heavy nuclei in the configuration-interaction shell model2023-12-31T21:41:37ZA low-energy enhancement (LEE) has been observed in the deexcitation $γ$-ray strength function ($γ$SF) of compound nuclei. The LEE has been a subject of intense experimental and theoretical interest since its discovery, and, if the LEE persists in heavy neutron-rich nuclei, it would have significant effects on calculations of r-process nucleosynthesis. Standard configuration-interaction (CI) shell-model calculations in medium-mass nuclei have attributed the LEE to the magnetic dipole $γ$SF but such calculations are computationally intractable in heavy nuclei. We review a combination of beyond-mean-field many-body methods within the framework of the CI shell model that enables the calculation of $γ$SF in heavy nuclei, and discuss the recent theoretical identification of a LEE in the magnetic dipole $γ$SF of lanthanide isotopes.2023-12-31T21:41:37Z8 pages, 5 figuresEPJ Web of Conferences 292, 01001 (2024)Y. AlhassidP. FantoA. Mercennehttp://arxiv.org/abs/2401.00546v3AllSpark: A Multimodal Spatio-Temporal General Intelligence Model with Ten Modalities via Language as a Reference Framework2025-01-07T13:31:01ZLeveraging multimodal data is an inherent requirement for comprehending geographic objects. However, due to the high heterogeneity in structure and semantics among various spatio-temporal modalities, the joint interpretation of multimodal spatio-temporal data has long been an extremely challenging problem. The primary challenge resides in striking a trade-off between the cohesion and autonomy of diverse modalities. This trade-off becomes progressively nonlinear as the number of modalities expands. Inspired by the human cognitive system and linguistic philosophy, where perceptual signals from the five senses converge into language, we introduce the Language as Reference Framework (LaRF), a fundamental principle for constructing a multimodal unified model. Building upon this, we propose AllSpark, a multimodal spatio-temporal general artificial intelligence model. Our model integrates ten different modalities into a unified framework. To achieve modal cohesion, AllSpark introduces a modal bridge and multimodal large language model (LLM) to map diverse modal features into the language feature space. To maintain modality autonomy, AllSpark uses modality-specific encoders to extract the tokens of various spatio-temporal modalities. Finally, observing a gap between the model's interpretability and downstream tasks, we designed modality-specific prompts and task heads, enhancing the model's generalization capability across specific tasks. Experiments indicate that the incorporation of language enables AllSpark to excel in few-shot classification tasks for RGB and point cloud modalities without additional training, surpassing baseline performance by up to 41.82\%. The source code is available at https://github.com/GeoX-Lab/AllSpark.2023-12-31T17:21:02Z19 pages, 19 tables, 3 figuresIEEE Transactions on Geoscience and Remote Sensing. 2025Run ShaoCheng YangQiujun LiQing ZhuYongjun ZhangYanSheng LiYu LiuYong TangDapeng LiuShizhong YangHaifeng Li10.1109/TGRS.2025.3526725http://arxiv.org/abs/2401.00553v6On Cohomology group of current Lie algebras2024-11-11T23:29:58ZIn this work we state a result that relates the cohomology groups of a Lie algebra $\mathfrak{g}$ and a current Lie algebra $\mathfrak{g} \otimes \mathcal{S}$, by means of a short exact sequence -- similar to the universal coefficients theorem for modules -- where $\mathcal{S}$ is a finite dimensional, commutative and associative algebra with unit over a field $\mathbb{F}$. Although this result can be applied to any Lie algebra, we determine the cohomology group of $\mathfrak{g} \otimes \mathcal{S}$, where $\mathfrak{g}$ is a semisimple Lie algebra.2023-12-31T17:46:45ZR. García-Delgadohttp://arxiv.org/abs/2401.00550v2Copies of Monomorphic Structures2024-06-06T09:10:50ZThe poset of copies of a relational structure ${\mathbb X}$ is the partial order $\langle {\mathbb P} ({\mathbb X}) ,\subset \rangle$, where ${\mathbb P} ({\mathbb X})=\{ Y\subset X: {\mathbb Y} \cong {\mathbb X}\}$. Investigating the classification of structures related to isomorphism of the Boolean completions ${\mathbb B}_{\mathbb X} ={\mathop{\rm ro}\nolimits}({\mathop{\rm sq}\nolimits} ({\mathbb P} ({\mathbb X}) ))$ we extend the results concerning linear orders to the class of structures definable in linear orders by first-order $Σ_0$-formulas (monomorphic structures). So, ${\mathbb B}_{\mathbb X} \cong {\mathbb B}_{\mathbb L}$ holds for some linear order ${\mathbb L}$, if ${\mathbb X}$ is definable in a $σ$-scattered (in particular, countable) or additively indecomposable linear order. For example, ${\mathbb B}_{\mathbb X} \cong {\mathop{\rm ro}\nolimits}({\mathbb S} )$, where ${\mathbb S}$ is the Sacks forcing, whenever ${\mathbb X}$ is a non-constant structure chainable by a real order type containing a perfect set.2023-12-31T17:40:32Z18 pagesMiloš S. Kurilićhttp://arxiv.org/abs/2401.00596v3Bulk medium properties of heavy-ion collisions from the beam energy scan with a multistage hydrodynamic model2024-07-08T16:37:17ZWe introduce a method to reconstruct full rapidity distributions of charged particle multiplicity and net proton yields, crucial for constraining the longitudinal dynamics of nuclear matter created in the beam energy scan program. Employing rapidity distributions within a multistage hydrodynamic model calibrated for Au+Au collisions at $\sqrt{s_\mathrm{NN}}=7.7-200\,$GeV, we estimate the total energy and baryon number deposited into the collision fireball, offering insights into initial dynamics and the identification of nuclear remnants. We explore the potential of rapidity-dependent measurements in probing equations of state at finite chemical potentials. Furthermore, we compare the freeze-out parameters derived from both hydrodynamics and thermal models, highlighting that the parameters extracted via thermal models represent averaged properties across rapidities.2023-12-31T22:22:44Zv1: 15 pages, 12 figures; v2: figs. 2 and 4 slightly adjusted; v3: Title slightly changed. A few short discussions were added, but there are no changes to the results or conclusions. Published in PRCPhys. Rev. C 110, 014904 (2024)Lipei Du10.1103/PhysRevC.110.014904http://arxiv.org/abs/2401.00544v2A Reliable Knowledge Processing Framework for Combustion Science using Foundation Models2024-01-02T03:03:18ZThis research explores the integration of large language models (LLMs) into scientific data assimilation, focusing on combustion science as a case study. Leveraging foundational models integrated with Retrieval-Augmented Generation (RAG) framework, the study introduces an approach to process diverse combustion research data, spanning experimental studies, simulations, and literature. The multifaceted nature of combustion research emphasizes the critical role of knowledge processing in navigating and extracting valuable information from a vast and diverse pool of sources. The developed approach minimizes computational and economic expenses while optimizing data privacy and accuracy. It incorporates prompt engineering and offline open-source LLMs, offering user autonomy in selecting base models. The study provides a thorough examination of text segmentation strategies, conducts comparative studies between LLMs, and explores various optimized prompts to demonstrate the effectiveness of the framework. By incorporating an external database, the framework outperforms a conventional LLM in generating accurate responses and constructing robust arguments. Additionally, the study delves into the investigation of optimized prompt templates for the purpose of efficient extraction of scientific literature. The research addresses concerns related to hallucinations and false research articles by introducing a custom workflow developed with a detection algorithm to filter out inaccuracies. Despite identified areas for improvement, the framework consistently delivers accurate domain-specific responses with minimal human oversight. The prompt-agnostic approach introduced holds promise for future deliberations. The study underscores the significance of integrating LLMs and knowledge processing techniques in scientific research, providing a foundation for advancements in data assimilation and utilization.2023-12-31T17:15:25Z38 pages and 10 figures; Fixed figure resolutionVansh SharmaVenkat Raman10.1016/j.egyai.2024.100365http://arxiv.org/abs/2401.00569v4Decision Making under Costly Sequential Information Acquisition: the Paradigm of Reversible and Irreversible Decisions2026-01-06T05:24:50ZDecision making in modern stochastic systems, including e-commerce platforms, financial markets and healthcare systems, has evolved into a multifaceted process that combines information acquisition and adaptive information sources. This paper initiates a study on such integrated settings, where these elements are not only fundamental but, also, interact in a complex and stochastically intertwined manner.
We introduce a relatively simple model, which, however, captures the involved novel elements. A decision maker (DM) may choose between an established product $A$ of known value and a new product $B$ whose value is unknown. In parallel, the DM observes signals about the unknown value of product $B$ and can, also, opt to exchange it for product $A$ if $B$ is initially chosen. Mathematically, the model gives rise to sequential optimal stopping problems with distinct informational regimes (before and after buying product $B$), differentiated by the initial, coarser signal and the subsequent, more accurate one. We analyze in detail the underlying problems using predominantly viscosity solution techniques, departing from the existing literature on information acquisition which is based on traditional optimal stopping arguments.
More broadly, the modeling approach introduced herein offers a novel framework for developing more complex interactions among decisions, information sources and information costs in stochastic environments, through a sequence of nested obstacle problems.2023-12-31T19:16:32ZRenyuan XuThaleia ZariphopoulouLuhao Zhanghttp://arxiv.org/abs/2401.00587v2Brain Tumor Segmentation Based on Deep Learning, Attention Mechanisms, and Energy-Based Uncertainty Prediction2024-03-14T19:02:51ZBrain tumors are one of the deadliest forms of cancer with a mortality rate of over 80%. A quick and accurate diagnosis is crucial to increase the chance of survival. However, in medical analysis, the manual annotation and segmentation of a brain tumor can be a complicated task. Multiple MRI modalities are typically analyzed as they provide unique information regarding the tumor regions. Although these MRI modalities are helpful for segmenting gliomas, they tend to increase overfitting and computation. This paper proposes a region of interest detection algorithm that is implemented during data preprocessing to locate salient features and remove extraneous MRI data. This decreases the input size, allowing for more aggressive data augmentations and deeper neural networks. Following the preprocessing of the MRI modalities, a fully convolutional autoencoder with soft attention segments the different brain MRIs. When these deep learning algorithms are implemented in practice, analysts and physicians cannot differentiate between accurate and inaccurate predictions. Subsequently, test time augmentations and an energy-based model were used for voxel-based uncertainty predictions. Experimentation was conducted on the BraTS benchmarks and achieved state-of-the-art segmentation performance. Additionally, qualitative results were used to assess the segmentation models and uncertainty predictions.2023-12-31T20:42:52Z11 pages, 6 figures, code available at https://github.com/WeToTheMoon/BrainTumorSegmentation, submitted to Computers in Biology and MedicineZachary SchwehrSriman Achantahttp://arxiv.org/abs/2401.00537v3Anisotropy of quadratic forms over global fields of characteristic $\neq$ 2 is diophantine2026-02-28T19:12:58ZWe prove that the set of anisotropic quadratic forms over global fields of characteristic different from 2 is a diophantine set. Our proof builds upon and extends the method of Koenigsmann, using tools from class field theory, the local-global principle, and advances on the diophantine definability of non-norm sets over global fields.2023-12-31T16:51:56ZGuang Huhttp://arxiv.org/abs/2401.00600v3Quasi-convergence of stability conditions2026-06-03T18:21:57ZWe develop a framework relating semiorthogonal decompositions of a triangulated category $\mathcal{C}$ to paths in its space of stability conditions. We prove that when $\mathcal{C}$ is the homotopy category of a smooth and proper idempotent complete pre-triangulated dg-category, every semiorthogonal decomposition whose factors admit a Bridgeland stability condition can be obtained from our framework.2023-12-31T22:41:34Z32 pages, 2 figures, formatting changes and typos corrected, isomorphic to publication version to appearDaniel Halpern-LeistnerJeffrey JiangAntonios-Alexandros Robotishttp://arxiv.org/abs/2401.00507v1Optimization of portfolios with cryptocurrencies: Markowitz and GARCH-Copula model approach2023-12-31T13:57:29ZThe growing interest in cryptocurrencies has drawn the attention of the financial world to this innovative medium of exchange. This study aims to explore the impact of cryptocurrencies on portfolio performance. We conduct our analysis retrospectively, assessing the performance achieved within a specific time frame by three distinct portfolios: one consisting solely of equities, bonds, and commodities; another composed exclusively of cryptocurrencies; and a third, which combines both 'traditional' assets and the best-performing cryptocurrency from the second portfolio.To achieve this, we employ the classic variance-covariance approach, utilizing the GARCH-Copula and GARCH-Vine Copula methods to calculate the risk structure. The optimal asset weights within the optimized portfolios are determined through the Markowitz optimization problem. Our analysis predominantly reveals that the portfolio comprising both cryptocurrency and traditional assets exhibits a higher Sharpe ratio from a retrospective viewpoint and demonstrates more stable performances from a prospective perspective. We also provide an explanation for our choice of portfolio optimization based on the Markowitz approach rather than CVaR and ES.2023-12-31T13:57:29ZVahidin JeleskovicClaudio LatiniZahid I. YounasMamdouh A. S. Al-Faryanhttp://arxiv.org/abs/2401.00514v1The mixing of two-pion and vector-meson states using staggered fermions2023-12-31T14:43:38ZIn this study we employ staggered fermions to calculate the two-pion taste singlet states at rest. Leveraging the Clebsch-Gordan coefficients of the symmetry group associated with staggered fermions, we effectively compute the $ππ$ contributions to the resting $ρ$-meson correlator. To discern the distinct energy states involved, we adopt a generalized eigenvalue problem-solving approach. This work will provide insight into the important role played by the two-pion contribution to the anomalous magnetic moment of the muon.
In this paper we present our group theoretic considerations and preliminary results on the contribution of two-pion states to the rho meson correlation function.2023-12-31T14:43:38Z8 pages, 2 tables, 5 figuresFabian J. FrechFinn M. StokesKalman K. SzaboBalint C. Tothhttp://arxiv.org/abs/2401.00516v1Kagomerization of transition metal monolayers induced by two-dimensional hexagonal boron nitride2023-12-31T15:05:45ZThe kagome lattice is an exciting solid state physics platform for the emergence of nontrivial quantum states driven by electronic correlations: topological effects, unconventional superconductivity, charge and spin density waves, and unusual magnetic states such as quantum spin liquids. While kagome lattices have been realized in complex multi-atomic bulk compounds, here we demonstrate from first-principles a process that we dub kagomerization, in which we fabricate a two-dimensional kagome lattice in monolayers of transition metals utilizing a hexagonal boron nitride (h-BN) overlayer. Surprisingly, h-BN induces a large rearrangement of the transition metal atoms supported on a fcc(111) heavy-metal surface. This reconstruction is found to be rather generic for this type of heterostructures and has a profound impact on the underlying magnetic properties, ultimately stabilizing various topological magnetic solitons such as skyrmions and bimerons. Our findings call for a reconsideration of h-BN as merely a passive capping layer, showing its potential for not only reconstructing the atomic structure of the underlying material, e.g. through the kagomerization of magnetic films, but also enabling electronic and magnetic phases that are highly sought for the next generation of device technologies.2023-12-31T15:05:45ZHangyu ZhouManuel dos Santos DiasYouguang ZhangWeisheng ZhaoSamir Lounishttp://arxiv.org/abs/2401.00518v1A Palm hierarchy for determinantal point processes with the confluent hypergeometric kernel, the decomposing measures in the problem of harmonic analysis on the infinite-dimensional unitary group2023-12-31T15:13:53ZThe main result of this note is that the shift of the parameter by 1 in the parameter space of decomposing measures in the problem of harmonic analysis on the infinite-dimensional unitary group corresponds to the taking of the reduced Palm measure at infinity for our decomposing measures. The proof proceeds by finite-dimensional approximation of our measures by orthogonal polynomial ensembles. The key remark is that the taking the reduced Palm measure commutes with the scaling limit transition from finite to infinite particle systems.2023-12-31T15:13:53Z23 pagesAlexander I. Bufetovhttp://arxiv.org/abs/2401.00519v1Multiplexed entanglement swapping with atomic-ensemble-based quantum memories in the single excitation regime2023-12-31T15:15:26ZEntanglement swapping (ES) between memory repeater links is critical for establishing quantum networks via quantum repeaters. So far, ES with atomic-ensemble-based memories has not been achieved. Here, we experimentally demonstrated ES between two entangled pairs of spin-wave memories via Duan-Lukin-Cirac-Zoller scheme. With a cloud of cold atoms inserted in a cavity, we produce non-classically-correlated spin-wave-photon pairs in 12 spatial modes and then prepare two entangled pairs of spin-wave memories via a multiplexed scheme. Via single-photon Bell measurement on retrieved fields from two memories, we project the two remaining memories never entangled previously into an entangled state with the measured concurrence of C = 0.0124(0.003). The successful probability of ES in our scheme is increased by three times, compared with that in non-multiplexed scheme. Our presented work shows that the generation of entanglement (C>0) between the remaining memory ensembles requires the average cross-correlation function of the spin-wave-photon pairs to be >30 .2023-12-31T15:15:26ZMinjie WangHaole JiaoJiajin LuWenxin FanShujing LiHai Wanghttp://arxiv.org/abs/2401.00535v1Actualised and future changes in regional economic growth through sea level rise2023-12-31T16:42:45ZThis study investigates the long-term economic impact of sea-level rise (SLR) on coastal regions in Europe, focusing on Gross Domestic Product (GDP). Using a novel dataset covering regional SLR and economic growth from 1900 to 2020, we quantify the relationships between SLR and regional GDP per capita across 79 coastal EU & UK regions. Our results reveal that the current SLR has already negatively influenced GDP of coastal regions, leading to a cumulative 4.7% loss at 39 cm of SLR. Over the 120 year period studied, the actualised impact of SLR on the annual growth rate is between -0.02% and 0.04%. Extrapolating these findings to future climate and socio-economic scenarios, we show that in the absence of additional adaptation measures, GDP losses by 2100 could range between -6.3% and -20.8% under the most extreme SLR scenario (SSP5-RCP8.5 High-end Ice, or -4.0% to -14.1% in SSP5-RCP8.5 High Ice). This statistical analysis utilising a century-long dataset, provides an empirical foundation for designing region-specific climate adaptation strategies to mitigate economic damages caused by SLR. Our evidence supports the argument for strategically relocating assets and establishing coastal setback zones when it is economically preferable and socially agreeable, given that protection investments have an economic impact.2023-12-31T16:42:45ZTheodoros ChatzivasileiadisIgnasi Cortes ArbuesJochen HinkelDaniel LinckeRichard S. J. Tolhttp://arxiv.org/abs/2401.00555v2The primordial black holes solution to the cosmological monopole problem2024-01-12T04:42:41ZRecently, the pulsar timing array (PTA) collaborations, including CPTA, EPTA, NANOGrav, and PPTA, announced that they detected a stochastic gravitational wave background spectrum in the nHz band. This may be relevant to the cosmological phase transition suggested by some models. Magnetic monopoles and primordial black holes (PBHs), two unsolved mysteries in the universe, may also have their production related to the cosmological phase transition. Inspired by that, we revisit the model proposed by Stojkovic and Freese, which involves PBHs accretion to solve the cosmological magnetic monopole problem. We further develop it by considering the increase in the mass of the PBHs during accretion and taking the effect of Hawking radiation into account. With these new considerations, we find that solutions to the problem still exist within a certain parameter space. In {addition}, we also generalize the analysis to PBHs with {an} extended distribution in mass. This may be a more interesting scenario because PBHs that have accreted magnetic monopoles might produce observable electromagnetic signals if they are massive enough to survive in the late universe.2023-12-31T18:04:03ZEur. Phys. J. C (2024) 84:31Xin-Zhe WangCan-Min Deng10.1140/epjc/s10052-024-12387-4http://arxiv.org/abs/2401.00568v1Extrapolation of Relative Treatment Effects using Change-point Survival Models2023-12-31T19:09:15ZIntroduction: Modelling of relative treatment effects is an important aspect to consider when extrapolating the long-term survival outcomes of treatments. Flexible parametric models offer the ability to accurately model the observed data, however, the extrapolated relative treatment effects and subsequent survival function may lack face validity. Methods: We investigate the ability of change-point survival models to estimate changes in the relative treatment effects, specifically treatment delay, loss of treatment effects and converging hazards. These models are implemented using standard Bayesian statistical software and propagate the uncertainty associate with all model parameters including the change-point location. A simulation study was conducted to assess the predictive performance of these models compared with other parametric survival models. Change-point survival models were applied to three datasets, two of which were used in previous health technology assessments. Results: Change-point survival models typically provided improved extrapolated survival predictions, particularly when the changes in relative treatment effects are large. When applied to the real world examples they provided good fit to the observed data while and in some situations produced more clinically plausible extrapolations than those generated by flexible spline models. Change-point models also provided support to a previously implemented modelling approach which was justified by visual inspection only and not goodness of fit to the observed data. Conclusions: We believe change-point survival models offer the ability to flexibly model observed data while also modelling and investigating clinically plausible scenarios with respect to the relative treatment effects.2023-12-31T19:09:15ZPhilip CooneyArthur Whitehttp://arxiv.org/abs/2401.00570v1On the classification of multiplicity-free Hamiltonian actions by regular proper symplectic groupoids2023-12-31T19:17:02ZIn this paper we study a natural generalization of symplectic toric manifolds in the context of regular Poisson manifolds of compact types. To be more precise, we consider a class of multiplicity-free Hamiltonian actions by regular proper symplectic groupoids that we call faithful. Given such a groupoid, we classify its faithful multiplicity-free Hamiltonian actions in terms of what we call Delzant subspaces of its orbit space -- certain `suborbifolds with corners' satisfying the Delzant condition relative to the integral affine orbifold structure of the orbit space. This encompasses both the classification of symplectic toric manifolds (due to Delzant) in terms of Delzant polytopes and the classification of proper Lagrangian fibrations over an integral affine base manifold (due to Duistermaat) in terms of a sheaf cohomology group. Each Delzant subspace comes with an orbifold version of this cohomology, the degree one part of which classifies faithful multiplicity-free Hamiltonian actions with momentum map image equal to the Delzant subspace, provided there exists such an action. The obstruction to existence is encoded by a degree two class in this cohomology: the Lagrangian Dixmier-Douady class. In addition to the above, we introduce another invariant, which leads to a variation of our classification result involving only classical sheaf cohomology and the group cohomology of certain modules for the isotropy groups of the groupoid.2023-12-31T19:17:02ZMaarten Molhttp://arxiv.org/abs/2401.00571v2Knot concordance, the point class in instanton homology and Donaldson invariants2024-02-24T01:23:50ZWe define an invariant ${\varphi}$ for knots in the 3-sphere by means of Donaldson invariants and Floer's instanton homology. Some basic properties of this invariant are established and it is shown that ${\varphi}$ coincides with a special case of an invariant defined by Froyshov2023-12-31T19:17:20ZRevision: previous version determined connect sum bound (Theorem 1(d)(e)) incorrectly. Corollaries and Examples updated. Section 4 revisedYuhan Limhttp://arxiv.org/abs/2401.00593v2Simplicity bias, algorithmic probability, and the random logistic map2024-04-08T23:32:29ZSimplicity bias is an intriguing phenomenon prevalent in various input-output maps, characterized by a preference for simpler, more regular, or symmetric outputs. Notably, these maps typically feature high-probability outputs with simple patterns, whereas complex patterns are exponentially less probable. This bias has been extensively examined and attributed to principles derived from algorithmic information theory and algorithmic probability. In a significant advancement, it has been demonstrated that the renowned logistic map and other one-dimensional maps exhibit simplicity bias when conceptualized as input-output systems. Building upon this work, our research delves into the manifestations of simplicity bias within the random logistic map, specifically focusing on scenarios involving additive noise.
We discover that simplicity bias is observable in the random logistic map for specific ranges of $μ$ and noise magnitudes. Additionally, we find that this bias persists even with the introduction of small measurement noise, though it diminishes as noise levels increase. Our studies also revisit the phenomenon of noise-induced chaos, particularly when $μ=3.83$, revealing its characteristics through complexity-probability plots. Intriguingly, we employ the logistic map to illustrate a paradoxical aspect of data analysis: more data adhering to a consistent trend can occasionally lead to \emph{reduced} confidence in extrapolation predictions, challenging conventional wisdom.
We propose that adopting a probability-complexity perspective in analyzing dynamical systems could significantly enrich statistical learning theories related to series prediction and analysis. This approach not only facilitates a deeper understanding of simplicity bias and its implications but also paves the way for novel methodologies in forecasting complex systems behavior.2023-12-31T22:08:34ZBoumediene HamziKamaludin Dingle10.13140/RG.2.2.29746.79048http://arxiv.org/abs/2401.00574v4Exact WKB analysis for ${\cal PT}$ symmetric quantum mechanics: Study of the Ai-Bender-Sarkar conjecture2024-03-29T03:42:52ZWe consider exact WKB analysis to a ${\cal PT}$ symmetric quantum mechanics defined by the potential, $V(x) = ω^2 x^2 + g x^2(i x)^{\varepsilon=2}$ with $ω\in {\mathbb R}_{\ge 0}$, $g \in {\mathbb R} _{> 0}$. We in particular aim to verify a conjecture proposed by Ai-Bender-Sarkar (ABS), that pertains to a relation between $D$-dimensional ${\cal PT}$-symmetric theories and analytic continuation (AC) of Hermitian theories concerning the energy spectrum or Euclidean partition function. For the purpose, we construct energy quantization conditions by exact WKB analysis and write down their transseries solution by solving the conditions. By performing alien calculus to the energy solutions, we verify validity of the ABS conjecture and seek a possibility of its alternative form by Borel resummation theory if it is violated. Our results claim that the validity of the ABS conjecture drastically changes depending on whether $ω> 0$ or $ω= 0$: If $ω>0$, then the ABS conjecture is violated when exceeding the semi-classical level of the first non-perturbative order, but its alternative form is constructable by Borel resummation theory. The ${\cal PT}$ and the AC energies are related to each other by a one-parameter Stokes automorphism, and a median resummed form, which corresponds to a formal exact solution, of the AC energy (resp. ${\cal PT}$ energy) is directly obtained by acting Borel resummation to a transseries solution of the ${\cal PT}$ energy (resp. AC energy). If $ω= 0$, then, with respect to the inverse energy level-expansion, not only perturbative/non-perturbative structures of the ${\cal PT}$ and the AC energies but also their perturbative parts do not match with each other. These energies are independent solutions, and no alternative form of the ABS conjecture can be reformulated by Borel resummation theory.2023-12-31T19:24:38Z40 pages, 13 figures, v2: typo and minor corrections, v3: typo and grammar corrections, references added, another example added in Sec. 3.1, structure slightly modified, v4: edit format is changed, minor revision, Sec VI A is modified, the conclusion is unchanged, accepted in PRDPhys. Rev. D 109, 085023 (2024)Syo Kamata10.1103/PhysRevD.109.085023http://arxiv.org/abs/2401.00522v2Analytical Critical Phenomena of Rotating Regular AdS Black Holes with Dark Energy2024-08-16T14:37:26ZThis study focuses on precisely calculating analytical critical points for rotating regular AdS black holes, examining scenarios with and without external dark field contributions. Importantly, it represents the first attempt to compute critical points for this specific class of rotating black holes. Our primary focus is on investigating the impact resulting from variations in the charge of nonlinear electrodynamics on the critical phenomena of rotating regular AdS black holes, while also incorporating the influence of quintessence field contributions. The analytical investigation is concentrated on the horizon radius, employing two distinct approaches to simplify the complexity and length of the calculations. Furthermore, our examination extends to deciphering the intricate relationship between dark energy and critical phenomena. This involves visually portraying a range of critical behaviors and detailing a recent discovery regarding how the intensity of quintessence affects phase transitions. The shift in these transitions conform to either a concave or convex function, a characteristic dependent on the sign of quintessence intensity.2023-12-31T15:25:26Z19 pages, 6 figures, 2 Tables. Accepted for publication in International Journal of Modern Physics A (2024)International Journal of Modern Physics A (2024)Hayat. LaassiriAhmed. DaassouRachid. Benbrik10.1142/S0217751X24500891http://arxiv.org/abs/2401.00505v3Higher-Order Cellular Automata Generated Symmetry-Protected Topological Phases and Detection Through Multi-Point Strange Correlators2024-08-08T15:09:36ZIn computer and system sciences, higher-order cellular automata (HOCA) are a type of cellular automata that evolve over multiple time steps and generate complex patterns, which have various applications such as secret sharing schemes, data compression, and image encryption. In this paper, we introduce HOCA to quantum many-body physics and construct a series of symmetry-protected topological (SPT) phases of matter, in which symmetries are supported on a great variety of subsystems embbeded in the SPT bulk. We call these phases HOCA-generated SPT (HGSPT) phases. Specifically, we show that HOCA can generate not only well-understood SPTs with symmetries supported on either regular (e.g., line-like subsystems in the 2D cluster model) or fractal subsystems, but also a large class of unexplored SPTs with symmetries supported on more choices of subsystems. One example is \textit{mixed-subsystem SPT} that has either fractal and line-like subsystem symmetries simultaneously or two distinct types of fractal symmetries simultaneously. Another example is \textit{chaotic-subsystem SPT} in which chaotic-looking symmetries are significantly different from and thus cannot reduce to fractal or regular subsystem symmetries. We also introduce a new notation system to characterize HGSPTs. We prove that all possible subsystem symmetries in square lattice can be locally simulated by an HOCA generated symmetry. As the usual two-point strange correlators are trivial in most HGSPTs, we find that the nontrivial SPT orders can be detected by what we call \textit{multi-point strange correlators}. We propose a universal procedure to design the spatial configuration of the multi-point strange correlators for a given HGSPT phase. Specifically, we find deep connections between multi-point strange correlators and the spurious topological entanglement entropy (STEE), both exhibiting long range behavior in SRE states.2023-12-31T13:56:20ZAccepted by PRX QuantumPRX Quantum 5, 030342 (2024)Jie-Yu ZhangMeng-Yuan LiPeng Ye10.1103/PRXQuantum.5.030342http://arxiv.org/abs/2401.00526v2Krylov Spread Complexity of Quantum-Walks2024-09-03T18:28:13ZGiven the recent advances in quantum technology, the complexity of quantum states is an important notion. The idea of the Krylov spread complexity has come into focus recently with the goal of capturing this in a quantitative way. The present paper sheds new light on the Krylov complexity measure by exploring it in the context of continuous-time quantum-walks on graphs. A close relationship between Krylov spread complexity and the concept of limiting-distributions for quantum-walks is established. Moreover, using a graph optimization algorithm, quantum-walk graphs are constructed that have minimal and maximal long-time average Krylov $\bar C$-complexity. This reveals an empirical upper bound for the $\bar C$-complexity as a function of Hilbert space dimension and an exact lower bound.2023-12-31T16:06:35Zv2: references added, typos correctedPhys. Rev. A 110, 032206 (2024)Bhilahari Jeevanesan10.1103/PhysRevA.110.032206http://arxiv.org/abs/2401.00585v1Harmonic curvature in dimension four2023-12-31T20:41:25ZWe provide a step towards classifying Riemannian four-manifolds in which the curvature tensor has zero divergence, or -- equivalently -- the Ricci tensor Ric satisfies the Codazzi equation. Every known compact manifold of this type belongs to one of five otherwise-familiar classes of examples. The main result consists in showing that, if such a manifold (not necessarily compact or even complete) lies outside of the five classes -- a non-vacuous assumption -- then, at all points of a dense open subset, Ric has four distinct eigenvalues, while suitable local coordinates simultaneously diagonalize Ric, the metric and, in a natural sense, also the curvature tensor. Furthermore, in a local orthonormal frame formed by Ricci eigenvectors, the connection form (or, curvature tensor) has just twelve (or, respectively, six) possibly-nonzero components, which together satisfy a specific system, not depending on the point, of homogeneous polynomial equations. A part of the classification problem is thus reduced to a question in real algebraic geometry.2023-12-31T20:41:25Z34 pagesJournal of the Korean Mathematical Society, vol. 62 (2025), no. 1, pp. 217-252Andrzej Derdzinski10.4134/JKMS.j240001http://arxiv.org/abs/2401.00510v3Smoothness Estimation for Whittle-Matérn Processes on Closed Riemannian Manifolds2025-06-05T06:14:36ZThe family of Matérn kernels are often used in spatial statistics, function approximation and Gaussian process methods in machine learning. One reason for their popularity is the presence of a smoothness parameter that controls, for example, optimal error bounds for kriging and posterior contraction rates in Gaussian process regression. On closed Riemannian manifolds, we show that the smoothness parameter can be consistently estimated from the maximizer(s) of the Gaussian likelihood when the underlying data are from point evaluations of a Gaussian process and, perhaps surprisingly, even when the data comprise evaluations of a non-Gaussian process. The points at which the process is observed need not have any particular spatial structure beyond quasi-uniformity. Our methods are based on results from approximation theory for the Sobolev scale of Hilbert spaces. Moreover, we generalize a well-known equivalence of measures phenomenon related to Matérn kernels to the non-Gaussian case by using Kakutani's theorem.2023-12-31T14:28:31ZStochastic Processes and their Applications 189:104685, 2025Moritz Korte-StapffToni KarvonenEric Moulines10.1016/j.spa.2025.104685http://arxiv.org/abs/2401.00588v2Fairness in Serving Large Language Models2024-06-05T06:43:16ZHigh-demand LLM inference services (e.g., ChatGPT and BARD) support a wide range of requests from short chat conversations to long document reading. To ensure that all client requests are processed fairly, most major LLM inference services have request rate limits, to ensure that no client can dominate the request queue. However, this rudimentary notion of fairness also results in under-utilization of the resources and poor client experience when there is spare capacity. While there is a rich literature on fair scheduling, serving LLMs presents new challenges due to their unpredictable request lengths and their unique batching characteristics on parallel accelerators. This paper introduces the definition of LLM serving fairness based on a cost function that accounts for the number of input and output tokens processed. To achieve fairness in serving, we propose a novel scheduling algorithm, the Virtual Token Counter (VTC), a fair scheduler based on the continuous batching mechanism. We prove a 2x tight upper bound on the service difference between two backlogged clients, adhering to the requirement of work-conserving. Through extensive experiments, we demonstrate the superior performance of VTC in ensuring fairness, especially in contrast to other baseline methods, which exhibit shortcomings under various conditions. The reproducible code is available at https://github.com/Ying1123/VTC-artifact2023-12-31T21:15:54ZYing ShengShiyi CaoDacheng LiBanghua ZhuZhuohan LiDanyang ZhuoJoseph E. GonzalezIon Stoicahttp://arxiv.org/abs/2401.00533v2Convergence of the complex block Jacobi methods under the generalized serial pivot strategies2024-07-11T10:12:52ZThe paper considers the convergence of the complex block Jacobi diagonalization methods under the large set of the generalized serial pivot strategies. The global convergence of the block methods for Hermitian, normal and $J$-Hermitian matrices is proven. In order to obtain the convergence results for the block methods that solve other eigenvalue problems, such as the generalized eigenvalue problem, we consider the convergence of a general block iterative process which uses the complex block Jacobi annihilators and operators.2023-12-31T16:30:29Z30 pages, 3 figuresLinear Algebra Appl. 699 (2024) 421-458Erna BegovicVjeran Hari10.1016/j.laa.2024.07.012http://arxiv.org/abs/2401.00538v1Electric Charging Effects on Insulating Surfaces in Cryogenic Liquids2023-12-31T16:54:52ZThis paper presents a new technique to study the adsorption and desorption of ions and electrons on insulating surfaces in the presence of strong electric fields in cryoliquids. The experimental design consists of a compact cryostat coupled with a sensitive electro-optical Kerr device to monitor the stability of the electric fields. The behavior of nitrogen and helium ions on a poly(methyl methacrylate) (PMMA) surface was compared to a PMMA surface coated with a mixture of deuterated polystyrene and deuterated polybutadiene. Ion accumulation and removal on these surfaces were unambiguously observed. Within the precision of the data, both surfaces behave similarly for the physisorbed ions. The setup was also used to measure the (quasi-)static dielectric constant of PMMA at T = 70 K. The impact of the ion adsorption on the search for a neutron permanent electric dipole moment in a cryogenic environment, like the nEDM@SNS experiment, is discussed.2023-12-31T16:54:52ZWolfgang KorschMark BroeringAshok TimsinaKent K. H. LeungJoshua AbneyDmitry BudkerBradley W. FilipponeJiachen HeSuman KanduMark McCreaMurchhana RoyChristopher SwankWeijun Yaohttp://arxiv.org/abs/2401.00563v3KernelGPT: Enhanced Kernel Fuzzing via Large Language Models2025-03-13T22:00:21ZBugs in operating system kernels can affect billions of devices and users all over the world. As a result, a large body of research has been focused on kernel fuzzing, i.e., automatically generating syscall (system call) sequences to detect potential kernel bugs or vulnerabilities. Kernel fuzzing aims to generate valid syscall sequences guided by syscall specifications that define both the syntax and semantics of syscalls. While there has been existing work trying to automate syscall specification generation, this remains largely manual work, and a large number of important syscalls are still uncovered.
In this paper, we propose KernelGPT, the first approach to automatically synthesizing syscall specifications via Large Language Models (LLMs) for enhanced kernel fuzzing. Our key insight is that LLMs have seen massive kernel code, documentation, and use cases during pre-training, and thus can automatically distill the necessary information for making valid syscalls. More specifically, KernelGPT leverages an iterative approach to automatically infer the specifications, and further debug and repair them based on the validation feedback. Our results demonstrate that KernelGPT can generate more new and valid specifications and achieve higher coverage than state-of-the-art techniques. So far, by using newly generated specifications, KernelGPT has already detected 24 new unique bugs in Linux kernel, with 12 fixed and 11 assigned with CVE numbers. Moreover, a number of specifications generated by KernelGPT have already been merged into the kernel fuzzer Syzkaller, following the request from its development team.2023-12-31T18:47:33ZASPLOS 2025Chenyuan YangZijie ZhaoLingming Zhang10.1145/3676641.3716022http://arxiv.org/abs/2401.00539v3On the implied volatility of Inverse options under stochastic volatility models2025-04-14T12:51:43ZIn this paper we study short-time behavior of the at-the-money implied volatility for Inverse European options with fixed strike price. The asset price is assumed to follow a general stochastic volatility process. Using techniques of the Malliavin calculus such as the anticipating It^o's formula we first compute the level of the implied volatility of the option when the maturity converges to zero. Then, we find a short maturity asymptotic formula for the skew of the implied volatility that depends on the roughness of the volatility model. We also show that our results extend easily to Quanto-Inverse options. We apply our general results to the SABR and fractional Bergomi models, and provide some numerical simulations that confirm the accurateness of the asymptotic formula for the skew. Finally, we provide an empirical application using Bitcoin options traded on Debirit to show how our theoretical formulas can be used to model real market data of such options.2023-12-31T17:02:04ZarXiv admin note: text overlap with arXiv:2308.15341, arXiv:2208.01353Elisa AlòsEulalia NualartMakar Pravosudhttp://arxiv.org/abs/2401.00552v3How Network Topology Affects the Strength of Dangerous Power Grid Perturbations2025-08-24T21:33:10ZReasonably large perturbations may push a power grid from its stable synchronous state into an undesirable state. Identifying vulnerabilities in power grids by studying power grid stability against such perturbations can aid in preventing future blackouts. Probabilistic stability quantifiers such as basin stability, which measures the asymptotic stability of a system, and survivability, which measures the transient stability of a system, have been commonly used to quantify the stability of nodes in a power grid. However, these quantifiers do not provide information about the strength of perturbations that destabilize the system. To measure the strength of perturbations beyond which the stability of the system gets compromised, we employ two probabilistic distance-based stability measures -- basin stability bound, which deals with a system's asymptotic behaviour, and survivability bound, a newly defined stability measure that deals with a system's transient behaviour. Using these stability quantifiers, we conduct a detailed study on the impact of network topology on the strength of dangerous power grid perturbations. In this, we uncover a new class of highly vulnerable nodes that were previously unknown. Additionally, we establish connections with tree-like network structures and node connectivity to lowly stable nodes.2023-12-31T17:44:03ZCalvin AlvaresSoumitro Banerjeehttp://arxiv.org/abs/2401.00512v2A parametricity-based formalization of semi-simplicial and semi-cubical sets2025-07-20T11:24:31ZSemi-simplicial and semi-cubical sets are commonly defined as presheaves over respectively, the semi-simplex or semi-cube category. Homotopy Type Theory then popularized an alternative definition, where the set of n-simplices or n-cubes are instead regrouped into the families of the fibers over their faces, leading to a characterization we call indexed. Moreover, it is known that semi-simplicial and semi-cubical sets are related to iterated Reynolds parametricity, respectively in its unary and binary variants. We exploit this correspondence to develop an original uniform indexed definition of both augmented semi-simplicial and semi-cubical sets, and fully formalize it in Coq.2023-12-31T14:39:00ZVersion corresponds to the published version (though with a different formatting). Associated formalization in Coq at https://github.com/artagnon/bonakMath. Struct. Comp. Sci. 35 (2025) e13Hugo HerbelinRamkumar Ramachandra10.1017/S096012952500009Xhttp://arxiv.org/abs/2401.00542v1A relaxation viewpoint to Unbalanced Optimal Transport: duality, optimality and Monge formulation2023-12-31T17:08:20ZWe present a general convex relaxation approach to study a wide class of Unbalanced Optimal Transport problems for finite non-negative measures with possibly different masses. These are obtained as the lower semicontinuous and convex envelope of a cost for non-negative Dirac masses. New general primal-dual formulations, optimality conditions, and metric-topological properties are carefully studied and discussed.2023-12-31T17:08:20Z57 pagesGiuseppe SavaréGiacomo Enrico Sodini