https://arxiv.org/api/8WpulSltl2J6tcXos0bpvpq99OY arXiv Query: search_query=&id_list=2401.00201,2401.00202,2401.00203,2401.00204,2401.00205,2401.00206,2401.00207,2401.00208,2401.00209,2401.00210,2401.00211,2401.00212,2401.00213,2401.00214,2401.00215,2401.00216,2401.00217,2401.00218,2401.00219,2401.00220,2401.00221,2401.00222,2401.00223,2401.00224,2401.00225,2401.00226,2401.00227,2401.00228,2401.00229,2401.00230,2401.00231,2401.00232,2401.00233,2401.00234,2401.00235,2401.00236,2401.00237,2401.00238,2401.00239,2401.00240,2401.00241,2401.00242,2401.00243,2401.00244,2401.00245,2401.00246,2401.00247,2401.00248,2401.00249,2401.00250,2401.00251,2401.00252,2401.00253,2401.00254,2401.00255,2401.00256,2401.00257,2401.00258,2401.00259,2401.00260,2401.00261,2401.00262,2401.00263,2401.00264,2401.00265,2401.00266,2401.00267,2401.00268,2401.00269,2401.00270,2401.00271,2401.00272,2401.00273,2401.00274,2401.00275,2401.00276,2401.00277,2401.00278,2401.00279,2401.00280,2401.00281,2401.00282,2401.00283,2401.00284,2401.00285,2401.00286,2401.00287,2401.00288,2401.00289,2401.00290,2401.00291,2401.00292,2401.00293,2401.00294,2401.00295,2401.00296,2401.00297,2401.00298,2401.00299,2401.00300&start=0&max_results=100 2026-07-02T22:41:31Z 100 100 0 http://arxiv.org/abs/2401.00216v2 Heavy Wilson Quarks and O($a$) Improvement: Nonperturbative Results for $b_{\rm g}$ 2024-02-01T06:38:09Z With Wilson quarks, on-shell O($a$) improvement of the lattice QCD action is achieved by including the Sheikholeslami-Wohlert term and two further operators of mass dimension 5, which amount to a mass-dependent rescaling of the bare parameters. We here focus on the rescaled bare coupling, $\tilde{g}_0^2 = g_0^2(1 + b_{\rm g} am_{\rm q})$, and the determination of $b_{\rm g}(g_0^2)$, which is currently only known to 1-loop order of perturbation theory. We derive suitable improvement conditions in the chiral limit and in a finite space-time volume and evaluate these for different gluonic observables, both with and without the gradient flow. The choice of $β$-values and the line of constant physics are motivated by the ALPHA collaboration's decoupling strategy to determine $α_s(m_Z)$. However, the improvement conditions and some insight into systematic effects may prove useful in other contexts, too. 2023-12-30T12:24:51Z 26 pages, 5 figures. Included additional references and fixed some typos. Matches published version J. High Energ. Phys. 2024, 188 (2024) Mattia Dalla Brida Roman Höllwieser Francesco Knechtli Tomasz Korzec Stefan Sint Rainer Sommer 10.1007/JHEP01(2024)188 http://arxiv.org/abs/2401.00225v1 Enhancing dysarthria speech feature representation with empirical mode decomposition and Walsh-Hadamard transform 2023-12-30T13:25:26Z Dysarthria speech contains the pathological characteristics of vocal tract and vocal fold, but so far, they have not yet been included in traditional acoustic feature sets. Moreover, the nonlinearity and non-stationarity of speech have been ignored. In this paper, we propose a feature enhancement algorithm for dysarthria speech called WHFEMD. It combines empirical mode decomposition (EMD) and fast Walsh-Hadamard transform (FWHT) to enhance features. With the proposed algorithm, the fast Fourier transform of the dysarthria speech is first performed and then followed by EMD to get intrinsic mode functions (IMFs). After that, FWHT is used to output new coefficients and to extract statistical features based on IMFs, power spectral density, and enhanced gammatone frequency cepstral coefficients. To evaluate the proposed approach, we conducted experiments on two public pathological speech databases including UA Speech and TORGO. The results show that our algorithm performed better than traditional features in classification. We achieved improvements of 13.8% (UA Speech) and 3.84% (TORGO), respectively. Furthermore, the incorporation of an imbalanced classification algorithm to address data imbalance has resulted in a 12.18% increase in recognition accuracy. This algorithm effectively addresses the challenges of the imbalanced dataset and non-linearity in dysarthric speech and simultaneously provides a robust representation of the local pathological features of the vocal folds and tracts. 2023-12-30T13:25:26Z Ting Zhu Shufei Duan Camille Dingam Huizhi Liang Wei Zhang http://arxiv.org/abs/2401.00227v1 Does the World Bank's Ease of Doing Business Index Matter for FDI? Findings from Africa 2023-12-30T13:27:31Z This paper investigates whether foreign investment (FDI) into Africa is at least partially responsive to World Bank-measured market friendliness. Specifically, I conducted analyses of four countries between 2009 and 2017, using cases that represent two of the highest scorers on the bank's Doing Business index as of 2008 (Mauritius and South Africa) and the two lowest scorers (DRC and CAR), and subsequently traced all four for growths or declines in FDI in relation to their scores in the index. The findings show that there is a moderate association between decreased costs of starting a business and growth of FDI. Mauritius, South Africa and the DRC reduced their total cost of starting a business by 71.7%, 143.7% and 122.9% for the entire period, and saw inward FDI increases of 167.6%, 79.8% and 152.21%, respectively. The CAR increased the cost of starting businesses but still saw increases in FDI. However, the country also saw the least amount of growth in FDI at only 13.3%. 2023-12-30T13:27:31Z Bhaso Ndzendze http://arxiv.org/abs/2401.00229v2 Thermal leptogenesis in nonextensive cosmology 2024-03-30T15:52:04Z Thermal leptogenesis is a mechanism that explains the observed asymmetry between matter and antimatter in the early universe. In this study, we review the impact of nonextensive Tsallis statistical mechanics on the early universe and study its effect on thermal leptogenesis. The study has found that the use of nonextensive statistical mechanics can affect the production of baryon asymmetry in thermal leptogenesis by modifying the equilibrium abundance of particles, decay, and washout parameters. Also, we show that nonextensive statistical mechanics potentially reduce the required right-handed neutrino mass scale. 2023-12-30T13:42:37Z 18 pages, 6 figures. arXiv admin note: text overlap with arXiv:2312.10677 Eur.Phys.J.C 84 (2024) 3, 340 Mehran Dehpour 10.1140/epjc/s10052-024-12697-7 http://arxiv.org/abs/2401.00230v2 Transformer Multivariate Forecasting: Less is More? 2024-03-07T10:25:20Z In the domain of multivariate forecasting, transformer models stand out as powerful apparatus, displaying exceptional capabilities in handling messy datasets from real-world contexts. However, the inherent complexity of these datasets, characterized by numerous variables and lengthy temporal sequences, poses challenges, including increased noise and extended model runtime. This paper focuses on reducing redundant information to elevate forecasting accuracy while optimizing runtime efficiency. We propose a novel transformer forecasting framework enhanced by Principal Component Analysis (PCA) to tackle this challenge. The framework is evaluated by five state-of-the-art (SOTA) models and four diverse real-world datasets. Our experimental results demonstrate the framework's ability to minimize prediction errors across all models and datasets while significantly reducing runtime. From the model perspective, one of the PCA-enhanced models: PCA+Crossformer, reduces mean square errors (MSE) by 33.3% and decreases runtime by 49.2% on average. From the dataset perspective, the framework delivers 14.3% MSE and 76.6% runtime reduction on Electricity datasets, as well as 4.8% MSE and 86.9% runtime reduction on Traffic datasets. This study aims to advance various SOTA models and enhance transformer-based time series forecasting for intricate data. Code is available at: https://github.com/jingjing-unilu/PCA_Transformer. 2023-12-30T13:44:23Z Jingjing Xu Caesar Wu Yuan-Fang Li Pascal Bouvry http://arxiv.org/abs/2401.00255v1 Adaptive Rank-based Tests for High Dimensional Mean Problems 2023-12-30T14:59:21Z The Wilcoxon signed-rank test and the Wilcoxon-Mann-Whitney test are commonly employed in one sample and two sample mean tests for one-dimensional hypothesis problems. For high-dimensional mean test problems, we calculate the asymptotic distribution of the maximum of rank statistics for each variable and suggest a max-type test. This max-type test is then merged with a sum-type test, based on their asymptotic independence offered by stationary and strong mixing assumptions. Our numerical studies reveal that this combined test demonstrates robustness and superiority over other methods, especially for heavy-tailed distributions. 2023-12-30T14:59:21Z Yu Zhang Long Feng http://arxiv.org/abs/2401.00279v1 Dimension of the singular set for $2$-valued stationary Lipschitz graphs 2023-12-30T16:53:59Z We prove that the singular set of a $2$-valued Lipschitz graph that is stationary for the area is of codimension $1$. 2023-12-30T16:53:59Z Comments are welcome! Jonas Hirsch Luca Spolaor http://arxiv.org/abs/2401.00292v1 A general framework for providing interval representations of Pareto optimal outcomes for large-scale bi- and tri-criteria MIP problems 2023-12-30T18:01:37Z The Multi-Objective Mixed-Integer Programming (MOMIP) problem is one of the most challenging. To derive its Pareto optimal solutions one can use the well-known Chebyshev scalarization and Mixed-Integer Programming (MIP) solvers. However, for a large-scale instance of the MOMIP problem, its scalarization may not be solved to optimality, even by state-of-the-art optimization packages, within the time limit imposed on the optimization. If a MIP solver cannot derive the optimal solution within the assumed time limit, it provides the optimality gap, which gauges the quality of the approximate solution. However, for the MOMIP case, no information is provided on the lower and upper bounds of the components of the Pareto optimal outcome. For the MOMIP problem with two and three objective functions, an algorithm is proposed to provide the so-called interval representation of the Pareto optimal outcome designated by the weighting vector when there is a time limit on solving the Chebyshev scalarization. Such interval representations can be used to navigate on the Pareto front. The results of several numerical experiments on selected large-scale instances of the multi-objective, multidimensional 0-1 knapsack problem illustrate the proposed approach. The limitations and possible enhancements of the proposed method are also discussed. 2023-12-30T18:01:37Z Grzegorz Filcek Janusz Miroforidis http://arxiv.org/abs/2401.00272v1 Dual-space Hierarchical Learning for Goal-guided Conversational Recommendation 2023-12-30T16:14:19Z Proactively and naturally guiding the dialog from the non-recommendation context (e.g., Chit-chat) to the recommendation scenario (e.g., Music) is crucial for the Conversational Recommender System (CRS). Prior studies mainly focus on planning the next dialog goal~(e.g., chat on a movie star) conditioned on the previous dialog. However, we find the dialog goals can be simultaneously observed at different levels, which can be utilized to improve CRS. In this paper, we propose Dual-space Hierarchical Learning (DHL) to leverage multi-level goal sequences and their hierarchical relationships for conversational recommendation. Specifically, we exploit multi-level goal sequences from both the representation space and the optimization space. In the representation space, we propose the hierarchical representation learning where a cross attention module derives mutually enhanced multi-level goal representations. In the optimization space, we devise the hierarchical weight learning to reweight lower-level goal sequences, and introduce bi-level optimization for stable update. Additionally, we propose a soft labeling strategy to guide optimization gradually. Experiments on two real-world datasets verify the effectiveness of our approach. Code and data are available here. 2023-12-30T16:14:19Z Accepted by Neurocomputing Can Chen Hao Liu Zeming Liu Xue Liu Dejing Dou http://arxiv.org/abs/2401.00265v1 An unconventional platform for two-dimensional Kagome flat bands on semiconductor surfaces 2023-12-30T15:34:20Z In condensed matter physics, the Kagome lattice and its inherent flat bands have attracted considerable attention for their potential to host a variety of exotic physical phenomena. Despite extensive efforts to fabricate thin films of Kagome materials aimed at modulating the flat bands through electrostatic gating or strain manipulation, progress has been limited. Here, we report the observation of a novel $d$-orbital hybridized Kagome-derived flat band in Ag/Si(111) $\sqrt{3}\times\sqrt{3}$ as revealed by angle-resolved photoemission spectroscopy. Our findings indicate that silver atoms on a silicon substrate form a Kagome-like structure, where a delicate balance in the hopping parameters of the in-plane $d$-orbitals leads to destructive interference, resulting in a flat band. These results not only introduce a new platform for Kagome physics but also illuminate the potential for integrating metal-semiconductor interfaces into Kagome-related research, thereby opening a new avenue for exploring ideal two-dimensional Kagome systems. 2023-12-30T15:34:20Z 7 pages, 4 figures Jae Hyuck Lee GwanWoo Kim Inkyung Song Yejin Kim Yeonjae Lee Sung Jong Yoo Deok-Yong Cho Jun-Won Rhim Jongkeun Jung Gunn Kim Changyoung Kim 10.1021/acsnano.4c05398 http://arxiv.org/abs/2401.00290v1 Red Teaming for Large Language Models At Scale: Tackling Hallucinations on Mathematics Tasks 2023-12-30T17:59:12Z We consider the problem of red teaming LLMs on elementary calculations and algebraic tasks to evaluate how various prompting techniques affect the quality of outputs. We present a framework to procedurally generate numerical questions and puzzles, and compare the results with and without the application of several red teaming techniques. Our findings suggest that even though structured reasoning and providing worked-out examples slow down the deterioration of the quality of answers, the gpt-3.5-turbo and gpt-4 models are not well suited for elementary calculations and reasoning tasks, also when being red teamed. 2023-12-30T17:59:12Z Accepted to The ART of Safety: Workshop on Adversarial testing and Red-Teaming for generative AI (IJCNLP-AACL 2023) Aleksander Buszydlik Karol Dobiczek Michał Teodor Okoń Konrad Skublicki Philip Lippmann Jie Yang http://arxiv.org/abs/2401.00220v3 Evidence of Grain Alignment by Magnetically Enhanced Radiative Torques from Multiwavelength Dust Polarization Modeling of HL Tau 2024-07-29T02:08:51Z Atacama Large Millimeter/Submillimeter Array (ALMA) has revolutionized the field of dust polarization in protoplanetary disks across multiple wavelengths. Previous observations and empirical modeling suggested multiple mechanisms of dust polarization toward HL Tau, including grain alignment and dust scattering. However, a detailed modeling of dust polarization based on grain alignment physics is not yet available. Here, using our updated POLARIS code, we perform numerical modeling of dust polarization arising from both grain alignment by Magnetically Enhanced Radiative Torque (MRAT) mechanism and self-scattering to reproduce the HL Tau polarization observed at three wavelengths 0.87, 1.3, and 3.1$\,$mm. Our modeling results show that the observed multi-wavelength polarization could be reproduced only when large grains contain embedded iron inclusions and those with slow internal relaxation must have wrong internal alignment (i.e., the grain's major axis parallel to its angular momentum). The abundance of iron embedded inside grains in the form of clusters is constrained to be $\gtrsim 16$%, and the number of iron atoms per cluster is $N_{\rm cl} \sim 9\times10^2$. Maximum grain sizes probed at wavelengths $λ$ = 0.87, 1.3, and 3.1$\,$mm are constrained at $\sim$ 60, 80, and 90$\,μ$m, respectively. 2023-12-30T13:09:11Z Published in ApJ Nguyen Tat Thang Pham Ngoc Diep Thiem Hoang Le Ngoc Tram Nguyen Bich Ngoc Nguyen Thi Phuong Bao Truong 10.3847/1538-4357/ad4f74 http://arxiv.org/abs/2401.00240v2 Evaluation of the Leggett-Garg inequality by means of the neutrino oscillations observed in reactor and accelerator experiments 2025-03-25T16:36:19Z We revisit the study of the violation of the Leggett-Garg inequality in neutrino oscillation data as a mean to test some of the fundamental aspects of quantum mechanics. In particular, we consider the results of the Daya Bay and RENO reactor experiments, and the MINOS and NOvA accelerator experiments. We find that DB and MINOS exhibit a strong manifestation of Leggett-Garg violation, whereas for RENO and NOvA data, the indication is weaker. Considering the particular baselines and energy ranges explored by each experiment, our results demonstrate that the Leggett-Garg violation is more evident for smaller baseline-to-energy ratios in all the data sets studied, a relevant aspect to consider when looking for evidence of quantum mechanical decoherence in neutrino oscillations. 2023-12-30T14:05:51Z 20 pages, 13 (16) figures. Comments are welcome. Performed revision of the calculations; minor changes in the results, and main conclusions unmodified Ricardo Zamora Barrios Mario A. Acero http://arxiv.org/abs/2401.00252v2 Cluster algebras and monotone Lagrangian tori 2025-08-06T07:09:03Z Motivated by the construction of Newton--Okounkov bodies and toric degenerations via cluster algebras in [GHKK18, FO25], we consider a family of Newton--Okounkov polytopes of a complex smooth Fano variety $X$ related by a composition of tropicalized cluster mutations. According to the work of [HK15], the toric degeneration associated with each Newton--Okounkov polytope $Δ$ in the family produces a completely integrable system of $X$ over $Δ$. We investigate circumstances in which each completely integrable system possesses a monotone Lagrangian torus fiber. We provide a sufficient condition, based on the data of tropical integer points and exchange matrices, for the family of constructed monotone Lagrangian tori to contain infinitely many monotone Lagrangian tori, no two of which are related by any symplectomorphism. By employing this criterion and exploiting the correspondence between the tropical integer points and the dual canonical basis elements, we generate infinitely many distinct monotone Lagrangian tori on flag manifolds of arbitrary type except in a few cases. 2023-12-30T14:38:34Z v2. 45 pages; incorporates referee's comments and suggestions Yunhyung Cho Myungho Kim Yoosik Kim Euiyong Park http://arxiv.org/abs/2401.00262v2 Finiteness conjecture for 3-manifolds obtained from handlebodies by attaching 2-handles 2024-12-23T16:42:05Z We study a generalized Witten's finiteness conjecture for the skein modules of oriented compact 3-manifolds with boundary. We formulate an equivalent version of the generalized finiteness conjecture using handlebodies and 2-handles, and prove the conjecture for some classes with the handlebodies of genus 2 and 3 using the equivalent version. 2023-12-30T15:29:23Z 21 pages, 13 figures, Added an example in Theorem 2.13 and Appendix A Algebr. Geom. Topol. 26 (2026) 1095-1114 Hiroaki Karuo Zhihao Wang 10.2140/agt.2026.26.1095 http://arxiv.org/abs/2401.00205v1 Consumer Manipulation via Online Behavioral Advertising 2023-12-30T11:10:35Z Online behavioral advertising (OBA) has a significant role in the digital economy. It allows advertisers to target consumers categorized according to their algorithmically inferred interests based on their behavioral data. As Alphabet and Meta gatekeep the Internet with their digital platforms and channel most of the consumer attention online, they are best placed to execute OBA and earn profits far exceeding fair estimations. There are increasing concerns that gatekeepers achieve such profitability at the expense of consumers, advertisers, and publishers who are dependent on their services to access the Internet. In particular, some claim that OBA systematically exploits consumers' decision-making vulnerabilities, creating internet infrastructure and relevant markets that optimize for consumer manipulation. Intuitively, consumer manipulation via OBA comes in tension with the ideal of consumer autonomy in liberal democracies. Nevertheless, academia has largely overlooked this phenomenon and instead has primarily focused on privacy and discrimination concerns of OBA. This article redirects academic discourse and regulatory focus on consumer manipulation via OBA. In doing so, first, this article elaborates on how OBA works. Second, it constructs an analytic framework for understanding manipulation. Third, it applies the theory of manipulation to OBA. As a result, this article illustrates the extent to which OBA leads to consumer manipulation. Crucially, this article is purely analytic and avoids normative evaluation of consumer manipulation via OBA. Evaluating consumer manipulation harms of OBA is an equally important but separate task and is pursued in another publication. 2023-12-30T11:10:35Z Lex Zard http://arxiv.org/abs/2401.00221v2 Structural Insights and an IP-based Solution Method for Patient-to-room Assignment under Consideration of Single Room Entitlements 2024-02-19T15:07:13Z Patient-to-room assignment (PRA) is a scheduling problem in decision support for hospitals. It consists of assigning patients to rooms according to certain objectives, e.g., avoiding transfers and respecting single-room requests. This work presents combinatorial insights about the feasibility of PRA and about the assignment of patients to single rooms. We further compare different IP-formulations for PRA as well as the influence of different objectivs on the runtime. Based on these results, we develop a fast IP-based solution approach which obtains high quality solution. The applicability is verified through a computational study with instances derived from real-world data. Results indicate that large, real world instances can be solved to a high degree of optimality within (fractions of) seconds. 2023-12-30T13:10:22Z Tabea Brandt Christina Büsing Felix Engelhardt http://arxiv.org/abs/2401.00232v2 A Review on Low-Dose Emission Tomography Post-Reconstruction Denoising with Neural Network Approaches 2024-01-15T15:35:23Z Low-dose emission tomography (ET) plays a crucial role in medical imaging, enabling the acquisition of functional information for various biological processes while minimizing the patient dose. However, the inherent randomness in the photon counting process is a source of noise which is amplified in low-dose ET. This review article provides an overview of existing post-processing techniques, with an emphasis on deep neural network (NN) approaches. Furthermore, we explore future directions in the field of NN-based low-dose ET. This comprehensive examination sheds light on the potential of deep learning in enhancing the quality and resolution of low-dose ET images, ultimately advancing the field of medical imaging. 2023-12-30T13:49:13Z 16 pages, 6 figures Alexandre Bousse Venkata Sai Sundar Kandarpa Kuangyu Shi Kuang Gong Jae Sung Lee Chi Liu Dimitris Visvikis 10.1109/TRPMS.2023.3349194 http://arxiv.org/abs/2401.00237v1 A Novel Approach for Defect Detection of Wind Turbine Blade Using Virtual Reality and Deep Learning 2023-12-30T13:58:50Z Wind turbines are subjected to continuous rotational stresses and unusual external forces such as storms, lightning, strikes by flying objects, etc., which may cause defects in turbine blades. Hence, it requires a periodical inspection to ensure proper functionality and avoid catastrophic failure. The task of inspection is challenging due to the remote location and inconvenient reachability by human inspection. Researchers used images with cropped defects from the wind turbine in the literature. They neglected possible background biases, which may hinder real-time and autonomous defect detection using aerial vehicles such as drones or others. To overcome such challenges, in this paper, we experiment with defect detection accuracy by having the defects with the background using a two-step deep-learning methodology. In the first step, we develop virtual models of wind turbines to synthesize the near-reality images for four types of common defects - cracks, leading edge erosion, bending, and light striking damage. The Unity perception package is used to generate wind turbine blade defects images with variations in background, randomness, camera angle, and light effects. In the second step, a customized U-Net architecture is trained to classify and segment the defect in turbine blades. The outcomes of U-Net architecture have been thoroughly tested and compared with 5-fold validation datasets. The proposed methodology provides reasonable defect detection accuracy, making it suitable for autonomous and remote inspection through aerial vehicles. 2023-12-30T13:58:50Z Md Fazle Rabbi Bill Solayman Hossain Emon Bill Ehtesham Mahmud Nishat Bill Tzu-Liang Bill Tseng Atira Ferdoushi Chun-Che Huang Md Fashiar Rahman http://arxiv.org/abs/2401.00253v2 The maximum sum of sizes of non-empty cross $t$-intersecting families 2024-02-07T05:39:04Z Let $[n]:=\lbrace 1,2,\ldots,n \rbrace$, and $M$ be a set of positive integers. Denote the family of all subsets of $[n]$ with sizes in $M$ by $\binom{\left[n\right]}{M}$. The non-empty families $\mathcal{A}\subseteq\binom{\left[n\right]}{R}$ and $\mathcal{B}\subseteq \binom{\left[n\right]}{S}$ are said to be cross $t$-intersecting if $|A\cap B|\geq t$ for all $A\in \mathcal{A}$ and $B\in \mathcal{B}$. In this paper, we determine the maximum sum of sizes of non-empty cross $t$-intersecting families, and characterize the extremal families. Similar result for finite vector spaces is also proved. 2023-12-30T14:44:05Z Shuang Li Dehai Liu Deping Song Tian Yao http://arxiv.org/abs/2401.00257v1 Assessing replication success via skeptical mixture priors 2023-12-30T15:11:02Z There is a growing interest in the analysis of replication studies of original findings across many disciplines. When testing a hypothesis for an effect size, two Bayesian approaches stand out for their principled use of the Bayes factor (BF), namely the replication BF and the skeptical BF. In particular, the latter BF is based on the skeptical prior, which represents the opinion of an investigator who is unconvinced by the original findings and wants to challenge them. We embrace the skeptical perspective, and elaborate a novel mixture prior which incorporates skepticism while at the same time controlling for prior-data conflict within the original data. Consistency properties of the resulting skeptical mixture BF are provided together with an extensive analysis of the main features of our proposal. Finally, we apply our methodology to data from the Social Sciences Replication Project. In particular we show that, for some case studies where prior-data conflict is an issue, our method uses a more realistic prior and leads to evidence-classification for replication success which differs from the standard skeptical approach. 2023-12-30T15:11:02Z 35 pages, 6 figures Guido Consonni Leonardo Egidi http://arxiv.org/abs/2401.00278v2 Intrinsic Shear and Galaxy Alignments: A Quantitative Study Using the TATT model 2024-01-04T22:25:18Z The intrinsic alignment (IA) of galaxies acts as a systematic effect in weak lensing measurements and tends to introduce biases. It mimics the gravitational lensing signal which makes it difficult to distinguish it from the true gravitational weak lensing effect. Hence, it is critical to account for the noise for correctly interpreting the results. This study aims at a quantitative analysis of IA using the Tidal Alignment and Tidal Torquing (TATT) model. We also investigate how the signals for shear and galaxy-galaxy lensing behave upon changing the parameters of the TATT model. The data for this study was prepared with a computational pipeline based on the Cocoa model to explore the parameter space of the intrinsic shape signal. Through this work, we identify that linear terms of the intrinsic shape signal are dominant in the case of GGL while the higher-order terms dictate the shear signal. 2023-12-30T16:52:37Z The paper is being withdrawn as it was originally an undergraduate research project and, upon reflection, it appears not sufficiently exhaustive for the intended academic discourse. The withdrawal is intended to maintain the high standards of research and scholarship in the field Abinash Das http://arxiv.org/abs/2401.00283v2 Near-Space Communications: the Last Piece of 6G Space-Air-Ground-Sea Integrated Network Puzzle 2024-03-04T16:22:58Z This article presents a comprehensive study on the emerging near-space communications (NS-COM) within the context of space-air-ground-sea integrated network (SAGSIN). Specifically, we firstly explore the recent technical developments of NS-COM, followed by the discussions about motivations behind integrating NS-COM into SAGSIN. To further demonstrate the necessity of NS-COM, a comparative analysis between the NS-COM network and other counterparts in SAGSIN is conducted, covering aspects of deployment, coverage, channel characteristics and unique problems of NS-COM network. Afterwards, the technical aspects of NS-COM, including channel modeling, random access, channel estimation, array-based beam management and joint network optimization, are examined in detail. Furthermore, we explore the potential applications of NS-COM, such as structural expansion in SAGSIN communication, civil aviation communication, remote and urgent communication, weather monitoring and carbon neutrality. Finally, some promising research avenues are identified, including stratospheric satellite (StratoSat) -to-ground direct links for mobile terminals, reconfigurable multiple-input multiple-output (MIMO) and holographic MIMO, federated learning in NS-COM networks, maritime communication, electromagnetic spectrum sensing and adversarial game, integrated sensing and communications, StratoSat-based radar detection and imaging, NS-COM assisted enhanced global navigation system, NS-COM assisted intelligent unmanned system and free space optical (FSO) communication. Overall, this paper highlights that the NS-COM plays an indispensable role in the SAGSIN puzzle, providing substantial performance and coverage enhancement to the traditional SAGSIN architecture. 2023-12-30T17:15:09Z 28 pages, 8 figures, 2 tables Hongshan Liu Tong Qin Zhen Gao Tianqi Mao Keke Ying Ziwei Wan Li Qiao Rui Na Zhongxiang Li Chun Hu Yikun Mei Tuan Li Guanghui Wen Lei Chen Zhonghuai Wu Ruiqi Liu Gaojie Chen Shuo Wang Dezhi Zheng http://arxiv.org/abs/2401.00286v1 Autonomous Threat Hunting: A Future Paradigm for AI-Driven Threat Intelligence 2023-12-30T17:36:08Z The evolution of cybersecurity has spurred the emergence of autonomous threat hunting as a pivotal paradigm in the realm of AI-driven threat intelligence. This review navigates through the intricate landscape of autonomous threat hunting, exploring its significance and pivotal role in fortifying cyber defense mechanisms. Delving into the amalgamation of artificial intelligence (AI) and traditional threat intelligence methodologies, this paper delineates the necessity and evolution of autonomous approaches in combating contemporary cyber threats. Through a comprehensive exploration of foundational AI-driven threat intelligence, the review accentuates the transformative influence of AI and machine learning on conventional threat intelligence practices. It elucidates the conceptual framework underpinning autonomous threat hunting, spotlighting its components, and the seamless integration of AI algorithms within threat hunting processes.. Insightful discussions on challenges encompassing scalability, interpretability, and ethical considerations in AI-driven models enrich the discourse. Moreover, through illuminating case studies and evaluations, this paper showcases real-world implementations, underscoring success stories and lessons learned by organizations adopting AI-driven threat intelligence. In conclusion, this review consolidates key insights, emphasizing the substantial implications of autonomous threat hunting for the future of cybersecurity. It underscores the significance of continual research and collaborative efforts in harnessing the potential of AI-driven approaches to fortify cyber defenses against evolving threats. 2023-12-30T17:36:08Z Siva Raja Sindiramutty http://arxiv.org/abs/2401.00271v1 HybridGait: A Benchmark for Spatial-Temporal Cloth-Changing Gait Recognition with Hybrid Explorations 2023-12-30T16:12:13Z Existing gait recognition benchmarks mostly include minor clothing variations in the laboratory environments, but lack persistent changes in appearance over time and space. In this paper, we propose the first in-the-wild benchmark CCGait for cloth-changing gait recognition, which incorporates diverse clothing changes, indoor and outdoor scenes, and multi-modal statistics over 92 days. To further address the coupling effect of clothing and viewpoint variations, we propose a hybrid approach HybridGait that exploits both temporal dynamics and the projected 2D information of 3D human meshes. Specifically, we introduce a Canonical Alignment Spatial-Temporal Transformer (CA-STT) module to encode human joint position-aware features, and fully exploit 3D dense priors via a Silhouette-guided Deformation with 3D-2D Appearance Projection (SilD) strategy. Our contributions are twofold: we provide a challenging benchmark CCGait that captures realistic appearance changes across an expanded and space, and we propose a hybrid framework HybridGait that outperforms prior works on CCGait and Gait3D benchmarks. Our project page is available at https://github.com/HCVLab/HybridGait. 2023-12-30T16:12:13Z Yilan Dong Chunlin Yu Ruiyang Ha Ye Shi Yuexin Ma Lan Xu Yanwei Fu Jingya Wang http://arxiv.org/abs/2401.00296v2 T-duality across non-extremal horizons 2024-09-26T10:52:19Z When applying T-duality to a generic, non-extreme Killing horizon, T-duality is spacelike on one side and timelike on the other. We show, using simple examples from four-dimensional Einstein-Maxwell theory, that the image of the horizon is a singularity which can be understood as an interface between two different T-dual theories and their solutions. Using an embedding into type-II string theory, we show that the singularity occurs when scalars reach the boundary of moduli space, resulting in a breakdown of the effective field theory due to the presence of tensionless strings. 2023-12-30T18:18:00Z 58 pages, 5 figures. Revised version: discussion of results has been expanded and improved JHEP 09 (2024) 116 Maxime Medevielle Thomas Mohaupt http://arxiv.org/abs/2401.00245v2 Alternative Approaches for Estimating Highest-Density Regions 2024-06-14T21:55:23Z Among the variety of statistical intervals, highest-density regions (HDRs) stand out for their ability to effectively summarize a distribution or sample, unveiling its distinctive and salient features. An HDR represents the minimum size set that satisfies a certain probability coverage, and current methods for their computation require knowledge or estimation of the underlying probability distribution or density $f$. In this work, we illustrate a broader framework for computing HDRs, which generalizes the classical density quantile method introduced in the seminal paper of Hyndman (1996). The framework is based on neighbourhood measures, i.e., measures that preserve the order induced in the sample by $f$, and include the density $f$ as a special case. We explore a number of suitable distance-based measures, such as the $k$-nearest neighborhood distance, and some probabilistic variants based on copula models. An extensive comparison is provided, showing the advantages of the copula-based strategy, especially in those scenarios that exhibit complex structures (e.g., multimodalities or particular dependencies). Finally, we discuss the practical implications of our findings for estimating HDRs in real-world applications. 2023-12-30T14:18:29Z Main paper: 26 pages (7 Figures, 1 Table); Supplementary Material: 36 pages International Statistical Review 2024 Nina Deliu Brunero Liseo 10.1111/insr.12592 http://arxiv.org/abs/2401.00233v2 Thermodiffusion unipolar electric generator 2024-11-18T17:22:57Z A model of a conducting cylinder with a radial temperature gradient which creates an electric field that increases with time in the surrounding vacuum is examined. The conditions under which this model functions are pointed out. An electric field is also generated when a magnetic field exists along the axis of the cylinder. This article discusses the interactions of the thermal flux, magnetic field, and charge distribution. Four models are considered with different conditions for the supply of electrons from a central source and the possibility either of capturing electrons inside the cylinder or their freely leaving it through the outer boundary. 2023-12-30T13:50:45Z 12 pages, 4 figures G. S. Bisnovatyi-Kogan M. V. Glushikhina 10.54503/0571-7132-2024.67.3-423 http://arxiv.org/abs/2401.00250v3 In-medium mass shift of two-flavored heavy mesons, $B_c$, $B^*_c$, $B_s$, $B^*_s$, $D_s$ and $D^*_s$ 2024-12-05T00:46:41Z For the first time, we estimate the in-medium mass shift of the two-flavored heavy mesons $B_c, B_c^*, B_s, B_s^*, D_s$ and $D_s^*$ in symmetric nuclear matter. The estimates are made by evaluating the lowest order one-loop self-energies. The enhanced excitations of intermediate state heavy-light mesons in symmetric nuclear matter are the origin of their negative mass shift. This negative mass shift may be regarded as a signature of partial restoration of chiral symmetry in an empirical sense because the origin of the negative mass shift in the study is not directly related to the chiral symmetry mechanism. Our results show that the magnitude of the mass shift for the $B_c$ meson ($\bar{b} c$ or $b \bar{c}$) is larger than those of the $η_c (\bar{c} c)$ and $η_b (\bar{b} b)$, different from a naive expectation that it would be in between them. While, that of the $B_c^*$ shows the in between of the $J/ψ$ and $Υ$. We observe that the lighter vector meson excitation in each meson self-energy gives a dominant contribution for the corresponding meson mass shift, $B_c, B_s,$ and $D_s$. 2023-12-30T14:36:11Z 24 pages, 21 figures (44 .eps files for figures), changes: unnecessary files removed, 1 figure corrected, 2 figures added Phys.Rev.D 110 (2024) 094045 G. N. Zeminiani S. L. P. G. Beres K. Tsushima 10.1103/PhysRevD.110.094045 http://arxiv.org/abs/2401.00204v6 Electromagnetic Radiation from Binary Stars Mediated by Ultralight Scalar 2025-05-15T11:07:14Z We present the electromagnetic (EM) dipole radiation flux from an eccentric Keplerian binary endowed with scalar charges, in the presence of scalar-photon coupling $φA_μA^μ$ or $φF_{μν}F^{μν}$. The scalar radiation is suppressed for orbital frequency below the scalar mass, while the scalar-mediated indirect EM radiation survives. We examine the constraints imposed on the scalar-photon and scalar-charge couplings by the current observational data of pulsar binaries, in case that the scalar charge is given by the muon number. The general extensions of the calculation to the quadrupole order and hyperbolic orbit are also discussed. 2023-12-30T11:07:57Z 30 pages, 5 figures, 2 tables Int.J.Mod.Phys.D 34 (2025) 05, 2550022 Ya-Ze Cheng Wen-Hao Wu Yan Cao 10.1142/S0218271825500221 http://arxiv.org/abs/2401.00274v1 Quadratic scaling path integral molecular dynamics for fictitious identical particles and its application to fermion systems 2023-12-30T16:27:03Z Recently, fictitious identical particles have provided a promising way to overcome the fermion sign problem and have been used in path integral Monte Carlo (PIMC) to accurately simulate warm dense matter with up to 1000 electrons (T. Dornheim et al., arXiv:2311.08098 (2023)). The inclusion of fictitious identical particles in path integral molecular dynamics (PIMD) can provide another way to simulate fermion systems. In a recent paper (J. Chem. Phys. 159, 154107 (2023)), Feldman and Hirshberg improved the recursive formula for PIMD of N identical bosons, significantly reducing the computational complexity from $O(PN^3)$ to $O(N^2+PN)$. In this paper, we extend this latest recursive formula for bosons to PIMD of fictitious identical particles to improve the efficiency of simulating fermion systems. We also provide the virial estimator for calculating energy by using the recursive technique. As an example, we use the quadratic scaling PIMD for fictitious identical particles to study the simulation of hundreds of fermions in a two-dimensional periodic potential, in the hope of providing a simulation tool for two-dimensional Fermi-Hubbard model and other strongly correlated fermion systems, such as the simulation of ultracold fermionic gases in optical lattices. 2023-12-30T16:27:03Z 30 pages, 8 figures Phys. Rev. E 110, 065303 (2025) Yunuo Xiong Shujuan Liu Hongwei Xiong 10.1103/PhysRevE.110.065303 http://arxiv.org/abs/2401.00247v1 Probing the Limits and Capabilities of Diffusion Models for the Anatomic Editing of Digital Twins 2023-12-30T14:21:30Z Numerical simulations can model the physical processes that govern cardiovascular device deployment. When such simulations incorporate digital twins; computational models of patient-specific anatomy, they can expedite and de-risk the device design process. Nonetheless, the exclusive use of patient-specific data constrains the anatomic variability which can be precisely or fully explored. In this study, we investigate the capacity of Latent Diffusion Models (LDMs) to edit digital twins to create anatomic variants, which we term digital siblings. Digital twins and their corresponding siblings can serve as the basis for comparative simulations, enabling the study of how subtle anatomic variations impact the simulated deployment of cardiovascular devices, as well as the augmentation of virtual cohorts for device assessment. However, while diffusion models have been characterized in their ability to edit natural images, their capacity to anatomically edit digital twins has yet to be studied. Using a case example centered on 3D digital twins of cardiac anatomy, we implement various methods for generating digital siblings and characterize them through morphological and topological analyses. We specifically edit digital twins to introduce anatomic variation at different spatial scales and within localized regions, demonstrating the existence of bias towards common anatomic features. We further show that such anatomic bias can be leveraged for virtual cohort augmentation through selective editing, partially alleviating issues related to dataset imbalance and lack of diversity. Our experimental framework thus delineates the limits and capabilities of using latent diffusion models in synthesizing anatomic variation for in silico trials. 2023-12-30T14:21:30Z 11 pages Karim Kadry Shreya Gupta Farhad R. Nezami Elazer R. Edelman 10.1038/s41746-024-01332-0 http://arxiv.org/abs/2401.00215v4 On families of elliptic curves $E_{p,q}:y^2=x^3-pqx$ that intersect the same line $L_{a,b}:y=\frac{a}{b}x$ of rational slope 2025-11-29T16:24:18Z Let $p$ and $q$ be two distinct odd primes, $p<q$ and $E_{p,q}:y^2=x^3-pqx$ be an elliptic curve. Fix a line $L_{a.b}:y=\frac{a}{b}x$ where $a\in \mathbb{Z},b\in \mathbb{N}$ and $(a,b)=1$. We study sufficient conditions that $p$ and $q$ must satisfy so that there are infinitely many elliptic curves $E_{p,q}$ that intersect $L_{a,b}$. 2023-12-30T12:24:19Z 24 pages, 9 figures Eldar Sultanow Anja Jeschke Amir Darwish Tfiha Madjid Tehrani William J Buchanan http://arxiv.org/abs/2401.00260v4 GazeCLIP: Enhancing Gaze Estimation Through Text-Guided Multimodal Learning 2025-03-08T13:37:22Z Visual gaze estimation, with its wide-ranging application scenarios, has garnered increasing attention within the research community. Although existing approaches infer gaze solely from image signals, recent advances in visual-language collaboration have demonstrated that the integration of linguistic information can significantly enhance performance across various visual tasks. Leveraging the remarkable transferability of large-scale Contrastive Language-Image Pre-training (CLIP) models, we address the open and urgent question of how to effectively apply linguistic cues to gaze estimation. In this work, we propose GazeCLIP, a novel gaze estimation framework that deeply explores text-face collaboration. Specifically, we introduce a meticulously designed linguistic description generator to produce text signals enriched with coarse directional cues. Furthermore, we present a CLIP-based backbone adept at characterizing text-face pairs for gaze estimation, complemented by a fine-grained multimodal fusion module that models the intricate interrelationships between heterogeneous inputs. Extensive experiments on three challenging datasets demonstrate the superiority of GazeCLIP, which achieves state-of-the-art accuracy. Our findings underscore the potential of using visual-language collaboration to advance gaze estimation and open new avenues for future research in multimodal learning for visual tasks. The implementation code and the pre-trained model will be made publicly available. 2023-12-30T15:24:50Z Jun Wang Hao Ruan Liangjian Wen Yong Dai Mingjie Wang http://arxiv.org/abs/2401.00264v5 Identification in Nonlinear Dynamic Panel Models under Partial Stationarity 2026-01-08T03:51:56Z This paper provides a general identification approach for a wide range of nonlinear panel data models, including binary choice, ordered response, and other types of limited dependent variable models. Our approach accommodates dynamic models with any number of lagged dependent variables as well as other types of endogenous covariates. Our identification strategy relies on a partial stationarity condition, which allows for not only an unknown distribution of errors, but also temporal dependencies in errors. We derive partial identification results under flexible model specifications and establish sharpness of our identified set in the binary choice setting. We demonstrate the robust finite-sample performance of our approach using Monte Carlo simulations, and apply the approach to the empirical analysis of income categories using various ordered choice models. 2023-12-30T15:33:10Z Wayne Yuan Gao Rui Wang http://arxiv.org/abs/2401.00201v1 The iterative conception of function and the iterative conception of set 2023-12-30T10:54:57Z Hilary Putnam once suggested that "the actual existence of sets as 'intangible objects' suffers... from a generalization of a problem first pointed out by Paul Benacerraf... are sets a kind of function or are functions a sort of set?" Sadly, he did not elaborate; my aim, here, is to do so on his behalf. There are well-known methods for treating sets as functions and functions as sets. But these do not raise any obvious philosophical or foundational puzzles. For that, we first need to provide a full-fledged function theory. I supply such a theory: it axiomatizes the iterative notion of function in exactly the same sense that ZF axiomatizes the iterative notion of set. Indeed, this function theory is synonymous with ZF. It might seem that set theory and function theory present us with rival foundations for mathematics, since they postulate different ontologies. But appearances are deceptive. Set theory and function theory provide the very same judicial foundation for mathematics. They do not supply rival metaphysical foundations; indeed, if they supply metaphysical foundations at all, then they supply the very same metaphysical foundations. 2023-12-30T10:54:57Z Tim Button http://arxiv.org/abs/2401.00209v1 AI and Tempo Estimation: A Review 2023-12-30T11:42:44Z The author's goal in this paper is to explore how artificial intelligence (AI) has been utilised to inform our understanding of and ability to estimate at scale a critical aspect of musical creativity - musical tempo. The central importance of tempo to musical creativity can be seen in how it is used to express specific emotions (Eerola and Vuoskoski 2013), suggest particular musical styles (Li and Chan 2011), influence perception of expression (Webster and Weir 2005) and mediate the urge to move one's body in time to the music (Burger et al. 2014). Traditional tempo estimation methods typically detect signal periodicities that reflect the underlying rhythmic structure of the music, often using some form of autocorrelation of the amplitude envelope (Lartillot and Toiviainen 2007). Recently, AI-based methods utilising convolutional or recurrent neural networks (CNNs, RNNs) on spectral representations of the audio signal have enjoyed significant improvements in accuracy (Aarabi and Peeters 2022). Common AI-based techniques include those based on probability (e.g., Bayesian approaches, hidden Markov models (HMM)), classification and statistical learning (e.g., support vector machines (SVM)), and artificial neural networks (ANNs) (e.g., self-organising maps (SOMs), CNNs, RNNs, deep learning (DL)). The aim here is to provide an overview of some of the more common AI-based tempo estimation algorithms and to shine a light on notable benefits and potential drawbacks of each. Limitations of AI in this field in general are also considered, as is the capacity for such methods to account for idiosyncrasies inherent in tempo perception, i.e., how well AI-based approaches are able to think and act like humans. 2023-12-30T11:42:44Z 9 pages Geoff Luck http://arxiv.org/abs/2401.00210v1 The Problem of Alignment 2023-12-30T11:44:59Z Large Language Models produce sequences learned as statistical patterns from large corpora. In order not to reproduce corpus biases, after initial training models must be aligned with human values, preferencing certain continuations over others. Alignment, which can be viewed as the superimposition of normative structure onto a statistical model, reveals a conflicted and complex interrelationship between language and technology. This relationship shapes theories of language, linguistic practice and subjectivity, which are especially relevant to the current sophistication in artificially produced text. We examine this practice of structuration as a two-way interaction between users and models by analysing how ChatGPT4 redacts perceived `anomalous' language in fragments of Joyce's Ulysses and the new linguistic practice of prompt engineering. We then situate this alignment problem historically, revisiting earlier postwar linguistic debates which counterposed two views of meaning: as discrete structures, and as continuous probability distributions. We discuss the largely occluded work of the Moscow Linguistic School, which sought to reconcile this opposition. Our attention to the Moscow School and later related arguments by Searle and Kristeva casts the problem of alignment in a new light: as one involving attention to the social structuration of linguistic practice, including structuration of anomalies that, like the Joycean text, exist in defiance of expressive conventions. These debates around the communicative orientation toward language can help explain some of the contemporary behaviours and interdependencies that take place between users and LLMs. 2023-12-30T11:44:59Z 23 pages, 1 figure Tsvetelina Hristova Liam Magee Karen Soldatic http://arxiv.org/abs/2401.00214v1 Evolutionary Dynamics with Randomly Distributed Benevolent Individuals 2023-12-30T12:22:05Z Understanding the evolution of cooperation is pivotal in biology and social science. Public resources sharing is a common scenario in the real world. In our study, we explore the evolutionary dynamics of cooperation on a regular graph with degree $k$, introducing the presence of a third strategy, namely the benevolence, who does not evolve over time, but provides a fixed benefit to all its neighbors. We find that the presence of the benevolence can foster the development of cooperative behavior and it follows a simple rule: $b/c > k - p_S(k-1)$. Our results provide new insights into the evolution of cooperation in structured populations. 2023-12-30T12:22:05Z Yuxin Geng Xingru Chen http://arxiv.org/abs/2401.00242v1 Laboratory Experiments of Model-based Reinforcement Learning for Adaptive Optics Control 2023-12-30T14:11:43Z Direct imaging of Earth-like exoplanets is one of the most prominent scientific drivers of the next generation of ground-based telescopes. Typically, Earth-like exoplanets are located at small angular separations from their host stars, making their detection difficult. Consequently, the adaptive optics (AO) system's control algorithm must be carefully designed to distinguish the exoplanet from the residual light produced by the host star. A new promising avenue of research to improve AO control builds on data-driven control methods such as Reinforcement Learning (RL). RL is an active branch of the machine learning research field, where control of a system is learned through interaction with the environment. Thus, RL can be seen as an automated approach to AO control, where its usage is entirely a turnkey operation. In particular, model-based reinforcement learning (MBRL) has been shown to cope with both temporal and misregistration errors. Similarly, it has been demonstrated to adapt to non-linear wavefront sensing while being efficient in training and execution. In this work, we implement and adapt an RL method called Policy Optimization for AO (PO4AO) to the GHOST test bench at ESO headquarters, where we demonstrate a strong performance of the method in a laboratory environment. Our implementation allows the training to be performed parallel to inference, which is crucial for on-sky operation. In particular, we study the predictive and self-calibrating aspects of the method. The new implementation on GHOST running PyTorch introduces only around 700 microseconds in addition to hardware, pipeline, and Python interface latency. We open-source well-documented code for the implementation and specify the requirements for the RTC pipeline. We also discuss the important hyperparameters of the method, the source of the latency, and the possible paths for a lower latency implementation. 2023-12-30T14:11:43Z Accepted for publication in JATIS Jalo Nousiainen Byron Engler Markus Kasper Chang Rajani Tapio Helin Cédric T. Heritier Sascha P. Quanz Adrian M. Glauser http://arxiv.org/abs/2401.00273v1 Investigating Zero-Shot Generalizability on Mandarin-English Code-Switched ASR and Speech-to-text Translation of Recent Foundation Models with Self-Supervision and Weak Supervision 2023-12-30T16:15:02Z This work evaluated several cutting-edge large-scale foundation models based on self-supervision or weak supervision, including SeamlessM4T, SeamlessM4T v2, and Whisper-large-v3, on three code-switched corpora. We found that self-supervised models can achieve performances close to the supervised model, indicating the effectiveness of multilingual self-supervised pre-training. We also observed that these models still have room for improvement as they kept making similar mistakes and had unsatisfactory performances on modeling intra-sentential code-switching. In addition, the validity of several variants of Whisper was explored, and we concluded that they remained effective in a code-switching scenario, and similar techniques for self-supervised models are worth studying to boost the performance of code-switched tasks. 2023-12-30T16:15:02Z Submitted to ICASSP 2024 Self-supervision in Audio, Speech and Beyond workshop Chih-Kai Yang Kuan-Po Huang Ke-Han Lu Chun-Yi Kuan Chi-Yuan Hsiao Hung-yi Lee http://arxiv.org/abs/2401.00276v1 Second-Order Uncertainty Quantification: Variance-Based Measures 2023-12-30T16:30:52Z Uncertainty quantification is a critical aspect of machine learning models, providing important insights into the reliability of predictions and aiding the decision-making process in real-world applications. This paper proposes a novel way to use variance-based measures to quantify uncertainty on the basis of second-order distributions in classification problems. A distinctive feature of the measures is the ability to reason about uncertainties on a class-based level, which is useful in situations where nuanced decision-making is required. Recalling some properties from the literature, we highlight that the variance-based measures satisfy important (axiomatic) properties. In addition to this axiomatic approach, we present empirical results showing the measures to be effective and competitive to commonly used entropy-based measures. 2023-12-30T16:30:52Z 22 pages, 10 figures Yusuf Sale Paul Hofman Lisa Wimmer Eyke Hüllermeier Thomas Nagler http://arxiv.org/abs/2401.00287v1 The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness 2023-12-30T17:37:06Z As Large Language Models (LLMs) play an increasingly pivotal role in natural language processing applications, their safety concerns become critical areas of NLP research. This paper presents Safety and Over-Defensiveness Evaluation (SODE) benchmark: a collection of diverse safe and unsafe prompts with carefully designed evaluation methods that facilitate systematic evaluation, comparison, and analysis over 'safety' and 'over-defensiveness.' With SODE, we study a variety of LLM defense strategies over multiple state-of-the-art LLMs, which reveals several interesting and important findings, such as (a) the widely popular 'self-checking' techniques indeed improve the safety against unsafe inputs, but this comes at the cost of extreme over-defensiveness on the safe inputs, (b) providing a safety instruction along with in-context exemplars (of both safe and unsafe inputs) consistently improves safety and also mitigates undue over-defensiveness of the models, (c) providing contextual knowledge easily breaks the safety guardrails and makes the models more vulnerable to generating unsafe responses. Overall, our work reveals numerous such critical findings that we believe will pave the way and facilitate further research in improving the safety of LLMs. 2023-12-30T17:37:06Z Neeraj Varshney Pavel Dolin Agastya Seth Chitta Baral http://arxiv.org/abs/2401.00291v2 Hub-collision avoidance and leaf-node options algorithm for fractal dimension and renormalization of complex networks 2024-01-04T07:23:34Z The box-covering method plays a fundamental role in the fractal property recognition and renormalization analysis of complex networks. This study proposes the hub-collision avoidance and leaf-node options (HALO) algorithm. In the box sampling process, a forward sampling rule (for avoiding hub collisions) and a reverse sampling rule (for preferentially selecting leaf nodes) are determined for bidirectional network traversal to reduce the randomness of sampling. In the box selection process, the larger necessary boxes are preferentially selected to join the solution by continuously removing small boxes. The compact-box-burning (CBB) algorithm, the maximum-excluded-mass-burning (MEMB) algorithm, the overlapping-box-covering (OBCA) algorithm, and the algorithm for combining small-box-removal strategy and maximum box sampling with a sampling density of 30 (SM30) are compared with HALO in experiments. Results on nine real networks show that HALO achieves the highest performance score and obtains 11.40%, 7.67%, 2.18%, and 8.19% fewer boxes than the compared algorithms, respectively. The algorithm determinism is significantly improved. The fractal dimensions estimated by covering four standard networks are more accurate. Moreover, different from MEMB or OBCA, HALO is not affected by the tightness of the hubs and exhibits a stable performance in different networks. Finally, the time complexities of HALO and the compared algorithms are all O(N^2), which is reasonable and acceptable. 2023-12-30T18:00:57Z 32 pages, 7 figures, conference or other essential info Chaos: An Interdisciplinary Journal of Nonlinear Science, 2022, 32(12): 123116 Feiyan Guo Jiajun Zhou Zhongyuan Ruan Jian Zhang Lin Qi 10.1063/5.0113001 http://arxiv.org/abs/2401.00293v1 Representation formulas for maximal monotone operators of type (D) in Banach spaces whose dual spaces are strictly convex 2023-12-30T18:03:46Z This work deals with a maximal monotone operator $A$ of type (D) in a Banach space whose dual space is strictly convex. We establish some representations for the value $Ax$ at a given point $x$ via its values at nearby points of $x$. We show that the faces of $Ax$ are contained in the set of all weak$^*$ convergent limits of bounded nets of the operator at nearby points of $x$, then we obtain a representation for $Ax$ by use of this set. In addition, representations for the support function of $Ax$ based on the minimal-norm selection of the operator in certain Banach spaces are given. 2023-12-30T18:03:46Z Comments are welcome! Nguyen B. Tran Tran N. Nguyen Huynh M. Hien http://arxiv.org/abs/2401.00219v3 Entanglement Entropy via Double Cone Regularization 2024-07-16T07:39:16Z This paper proposes an alternative regularization method for handling the ultraviolet behavior of entanglement entropy. Utilizing an $iε$ prescription in the Euclidean double cone geometry, it accurately reproduces the universal behavior of entanglement entropy. The method is demonstrated in the free boson theory in arbitrary dimensions and two-dimensional conformal field theories. The findings highlight the effectiveness of the $iε$ regularization method in addressing ultraviolet issues in quantum field theory and gravity, suggesting potential applications to other calculable quantities. 2023-12-30T13:07:26Z 5 pages and 1 figure. Added acknowledgement for v2, Published version v3 including minor corrections and improvements Taishi Kawamoto Yu-ki Suzuki http://arxiv.org/abs/2401.00294v2 Robust Quantum Control in Closed and Open Systems: Theory and Practice 2024-07-27T17:55:20Z Robust control of quantum systems is an increasingly relevant field of study amidst the second quantum revolution, but there remains a gap between taming quantum physics and robust control in its modern analytical form that culminated in fundamental performance bounds. With certain exceptions such as quantum optical systems that can be modeled as linear stochastic differential equations, quantum systems are not amenable to linear, time-invariant, measurement-based robust control techniques, and thus novel gap-bridging techniques must be developed. This survey is written for control theorists to provide a review of the current state of quantum control and outline the challenges faced in trying to apply modern robust control to quantum systems. We present issues that arise when applying classical robust control theory to quantum systems, typical methods used by quantum physicists to explore such systems and their robustness, as well as a discussion of open problems to be addressed in the field. We focus on general, practical applications and recent work to enable control researchers to contribute to advancing this burgeoning field. 2023-12-30T18:08:43Z Minor revisions to earlier version C. A. Weidner E. A. Reed J. Monroe B. Sheller S. O'Neil E. Maas E. A. Jonckheere F. C. Langbein S. G. Schirmer http://arxiv.org/abs/2401.00203v1 Coarsening of topological defects in 2D polar active matter 2023-12-30T11:05:54Z We numerically study the coarsening of topological defects in 2D polar active matter and make several interesting observations and predictions. (i) The long time state is characterized by nonzero density of defects, in stark contrast to theoretical expectations. (ii) The kinetics of defect coarsening shows power law decay to steady state, as opposed to exponential decay in thermal equilibrium. (iii) Observations (i) and (ii) together suggest emergent screening of topological charges due to activity. (iv) Nontrivial defect coarsening in the active model leads to nontrivial steady state patterns. We investigate, characterize, and validate these patterns and discuss their biological significance. 2023-12-30T11:05:54Z Movies are available in the ancillary folder Soft Matter, 2025,21, 77-86 Soumyadeep Mondal Pankaj Popli Sumantra Sarkar 10.1039/D4SM00788C http://arxiv.org/abs/2401.00295v2 Imperfect Entangling Power of Quantum Gates 2025-07-30T13:23:20Z Achieving perfect control over the parameters defining a quantum gate is, in general, a very challenging task, and at the same time, environmental interactions can introduce disturbances to the initial states as well. Here we address the problem of how the imperfections in unitaries and noise present in the input states affect the entanglement-generating power of a given quantum gate -- we refer to it as imperfect (noisy) entangling power. We observe that, when the parameters of a given unitary are chosen randomly from a Gaussian distribution centered around the desired mean, the quenched average entangling power -- averaged across multiple random samplings -- exhibits intriguing behavior like it may increase or show nonmonotonic behavior with the increase of disorder strength for certain classes of diagonal unitary operators. For arbitrary unitary operators, the quenched average power tends to stabilize, showing almost constant behavior with variation in the parameters instead of oscillating. Our observations also reveal that, in the presence of a local noise model, the input states that maximize the entangling power of a given unitary operator differ considerably from the noiseless scenario. Additionally, we report that the rankings among unitary operators according to their entangling power in the noiseless case change depending on the noise model and noise strength. 2023-12-30T18:13:26Z 17 pages, 16 figures Phys. Rev. A 112, 012417 (2025) Sudipta Mondal Samir Kumar Hazra Aditi Sen De 10.1103/jl2b-bxfn http://arxiv.org/abs/2401.00212v4 Physics-Informed Multi-Agent Reinforcement Learning for Distributed Multi-Robot Problems 2025-06-22T18:00:14Z The networked nature of multi-robot systems presents challenges in the context of multi-agent reinforcement learning. Centralized control policies do not scale with increasing numbers of robots, whereas independent control policies do not exploit the information provided by other robots, exhibiting poor performance in cooperative-competitive tasks. In this work we propose a physics-informed reinforcement learning approach able to learn distributed multi-robot control policies that are both scalable and make use of all the available information to each robot. Our approach has three key characteristics. First, it imposes a port-Hamiltonian structure on the policy representation, respecting energy conservation properties of physical robot systems and the networked nature of robot team interactions. Second, it uses self-attention to ensure a sparse policy representation able to handle time-varying information at each robot from the interaction graph. Third, we present a soft actor-critic reinforcement learning algorithm parameterized by our self-attention port-Hamiltonian control policy, which accounts for the correlation among robots during training while overcoming the need of value function factorization. Extensive simulations in different multi-robot scenarios demonstrate the success of the proposed approach, surpassing previous multi-robot reinforcement learning solutions in scalability, while achieving similar or superior performance (with averaged cumulative reward up to x2 greater than the state-of-the-art with robot teams x6 larger than the number of robots at training time). We also validate our approach on multiple real robots in the Georgia Tech Robotarium under imperfect communication, demonstrating zero-shot sim-to-real transfer and scalability across number of robots. 2023-12-30T12:12:35Z Paper accepted and published at IEEE T-RO Eduardo Sebastian Thai Duong Nikolay Atanasov Eduardo Montijano Carlos Sagues http://arxiv.org/abs/2401.00236v4 Inverse problems for elastic wave from Partial Cauchy Data: Uniqueness and Co-inversion for Shape and Impedance Function 2025-06-25T12:58:50Z We consider an inverse problem for the elastic wave of simultaneously reconstructing the impedance and the geometric information of the bounded body that is occupied by a homogeneous and isotropic elastic medium from the measured Cauchy data. A two-stage reconstruction method is proposed to realize simultaneous reconstruction of multiple targets. In the first step, we restore the aperture information by utilizing the observed Cauchy data that is measured on an accessible part of the boundary. In the second step, we start with the boundary condition and propose a novel iterative method to simultaneously reconstruct the missing boundary and the impedance function. Theoretically, we establish the uniqueness result of the co-inversion problem based on analyzing the properties of the corresponding operators. An explicit derivative is computed for the iterative method. Numerical examples are presented to test the effectiveness and efficiency of the proposed method. 2023-12-30T13:54:32Z Yao Sun Yan Chang Yukun Guo http://arxiv.org/abs/2401.00226v2 The Intrinsic Energy Resolution of LaBr$_3$(Ce) Crystal for GECAM 2025-01-03T13:15:41Z This study aims to provide an accurate estimation of the intrinsic resolution of LaBr$_3$(Ce) crystal through a combination of experimental and simulation methods. We re-analyzed the data from previous Wide-Angle Compton Coincidence (WACC) and Hard X-ray Calibration Facility (HXCF) experiments, conducted PMT Single-Photoelectron Calibration (SPEC) and radial non-uniformity (also called Spot Scanning, SS) experiments to acquire new data, and combined these results with Geant4 simulations to isolate the contribution of each physical process to the total energy resolution, thereby allowing for a precise estimation of the scintillator's intrinsic resolution. For 100 keV X-rays, the total energy resolution of LaBr$_3$(Ce) crystal is 3.99% $\pm$ 0.04% (expressed as 1-$σ$), with statistical fluctuations and intrinsic resolution as the main components, contributing 2.47% $\pm$ 0.00% and 3.06% $\pm$ 0.06%, respectively. We identify two main sources of intrinsic resolution: one primarily due to non-proportional scintillation, contributing 2.28% $\pm$ 0.00%, and the other due to fluctuations in the energy transfer process, contributing 2.04% $\pm$ 0.08%. We quantified six components of the total energy resolution and reconstructed the photon response using Geant4. The consistency between the reconstructed relative light yield and the experimental measurements validated the mass model of the LaBr$_3$(Ce) detector used in the simulations. 2023-12-30T13:26:14Z 20 pages, 14 figures Pei-Yi Feng Key Laboratory of Particle Astrophysics, Institute of High Energy Physics, Chinese Academy of Sciences, Beijing, China University of Chinese Academy of Sciences, Chinese Academy of Sciences, Beijing, China Xi-Lei Sun State Key Laboratory of Particle Detection and Electronics, Institute of High Energy Physics, Chinese Academy of Sciences, Beijing, China Zheng-Hua An Key Laboratory of Particle Astrophysics, Institute of High Energy Physics, Chinese Academy of Sciences, Beijing, China Cheng-Er Wang National Engineering Research Center for Rare Earth, Grirem Advanced Materials Co., Ltd., Beijing, China Da-Li Zhang Key Laboratory of Particle Astrophysics, Institute of High Energy Physics, Chinese Academy of Sciences, Beijing, China Xin-Qiao Li Key Laboratory of Particle Astrophysics, Institute of High Energy Physics, Chinese Academy of Sciences, Beijing, China Chao Zheng Key Laboratory of Particle Astrophysics, Institute of High Energy Physics, Chinese Academy of Sciences, Beijing, China University of Chinese Academy of Sciences, Chinese Academy of Sciences, Beijing, China Shao-Lin Xiong Key Laboratory of Particle Astrophysics, Institute of High Energy Physics, Chinese Academy of Sciences, Beijing, China Hong Lu Key Laboratory of Particle Astrophysics, Institute of High Energy Physics, Chinese Academy of Sciences, Beijing, China http://arxiv.org/abs/2401.00254v4 Morphing Tokens Draw Strong Masked Image Models 2025-03-21T09:24:14Z Masked image modeling (MIM) has emerged as a promising approach for pre-training Vision Transformers (ViTs). MIMs predict masked tokens token-wise to recover target signals that are tokenized from images or generated by pre-trained models like vision-language models. While using tokenizers or pre-trained models is viable, they often offer spatially inconsistent supervision even for neighboring tokens, hindering models from learning discriminative representations. Our pilot study identifies spatial inconsistency in supervisory signals and suggests that addressing it can improve representation learning. Building upon this insight, we introduce Dynamic Token Morphing (DTM), a novel method that dynamically aggregates tokens while preserving context to generate contextualized targets, thereby likely reducing spatial inconsistency. DTM is compatible with various SSL frameworks; we showcase significantly improved MIM results, barely introducing extra training costs. Our method facilitates MIM training by using more spatially consistent targets, resulting in improved training trends as evidenced by lower losses. Experiments on ImageNet-1K and ADE20K demonstrate DTM's superiority, which surpasses complex state-of-the-art MIM methods. Furthermore, the evaluation of transfer learning on downstream tasks like iNaturalist, along with extensive empirical studies, supports DTM's effectiveness. 2023-12-30T14:53:09Z 24 pages, 16 tables, 8 figures. To be presented at ICLR'25 Taekyung Kim Byeongho Heo Dongyoon Han http://arxiv.org/abs/2401.00208v1 Inpaint4DNeRF: Promptable Spatio-Temporal NeRF Inpainting with Generative Diffusion Models 2023-12-30T11:26:55Z Current Neural Radiance Fields (NeRF) can generate photorealistic novel views. For editing 3D scenes represented by NeRF, with the advent of generative models, this paper proposes Inpaint4DNeRF to capitalize on state-of-the-art stable diffusion models (e.g., ControlNet) for direct generation of the underlying completed background content, regardless of static or dynamic. The key advantages of this generative approach for NeRF inpainting are twofold. First, after rough mask propagation, to complete or fill in previously occluded content, we can individually generate a small subset of completed images with plausible content, called seed images, from which simple 3D geometry proxies can be derived. Second and the remaining problem is thus 3D multiview consistency among all completed images, now guided by the seed images and their 3D proxies. Without other bells and whistles, our generative Inpaint4DNeRF baseline framework is general which can be readily extended to 4D dynamic NeRFs, where temporal consistency can be naturally handled in a similar way as our multiview consistency. 2023-12-30T11:26:55Z Han Jiang Haosen Sun Ruoxuan Li Chi-Keung Tang Yu-Wing Tai http://arxiv.org/abs/2401.00217v1 Discretization-Based Solution Approaches for the Circle Packing Problem 2023-12-30T12:47:23Z The problem of packing a set of circles into the smallest surrounding container is considered. This problem arises in different application areas such as automobile, textile, food, and chemical industries. The so-called circle packing problem can be cast as a nonconvex quadratically constrained program, and is difficult to solve in general. An iterative solution approach based on a bisection-type algorithm on the radius of the larger circle is provided. The present algorithm discretizes the container into small cells and solves two different integer linear programming formulations proposed for a restricted and a relaxed version of the original problem. The present algorithm is enhanced with solution space reduction, bound tightening and variable elimination techniques. Then, a computational study is performed to evaluate the performance of the algorithm. The present algorithm is compared with BARON and Gurobi that solve the original nonlinear formulation and heuristic methods from literature, and obtain promising results. 2023-12-30T12:47:23Z Rabia Taşpınar Burak Kocuk http://arxiv.org/abs/2401.00223v1 On Performance of Integrated Satellite HAPS Ground Communication: Aerial IRS Node vs Terrestrial IRS Node 2023-12-30T13:19:18Z With a motive of ubiquitous connectivity over the globe with enhanced spectral efficiency, intelligent reflecting surfaces (IRS) integrated satellite-terrestrial communications is a topic of research interest in an infrastructure-deficient remote terrains. In line with this vision, this paper entails the performance analysis of satellite-terrestrial networks leveraging both aerial and terrestrial IRS nodes, with the support of high altitude platforms over diverse fading channels including shadowed Rician, Rician, and Nakagami-$m$ fading channels. The merits of IRS in enhancing spectral efficiency is analyzed through closed-form expressions of outage probability and ergodic rate. Further, the average symbol error rate analysis for the higher-order quadrature amplitude modulation (QAM) schemes such as hexagonal QAM, rectangular QAM, cross QAM, and square QAM is performed. Practical constraints like antenna gains, path loss, and link fading are considered to characterize the satellite terrestrial links. Finally, a comparison between the high-altitude platforms based IRS node and terrestrial IRS nodes is performed and various insights are drawn under various fading scenarios and path loss conditions. This paper contribute towards understanding and potential implementation of IRS-integrated satellite-terrestrial networks for efficient and reliable communication. 2023-12-30T13:19:18Z Parvez Shaik Kamal Kishore Garg Praveen Kumar Singya Vimal Bhatia Mohamed-Slim Alouini http://arxiv.org/abs/2401.00224v1 Edge and bulk states in Weyl-orbit quantum Hall effect as studied by Corbino measurements 2023-12-30T13:24:24Z We investigate edge and bulk states in Weyl-orbit based quantum Hall effect by measuring a Corbino-type device fabricated from a topological Dirac semimetal (Cd1-xZnx)3As2 film. Clear quantum Hall plateaus are observed when measuring one-sided terminals of the Corbino-type device. This indicates that edge states of the Weyl-orbit quantum Hall effect form closed trajectories consisting of Fermi arcs and chiral zero modes independently on inner and outer sides. On the other hand, the bulk resistance does not diverge at fields where the quantum Hall plateau appears, suggesting that the Weyl orbits in the bulk region are not completely localized when applying electric current through the bulk region. 2023-12-30T13:24:24Z Journal of the Physical Society of Japan 93, 023706 (2024) Yusuke Nakazawa Ryosuke Kurihara Masatoshi Miyazawa Shinichi Nishihaya Markus Kriener Masashi Tokunaga Masashi Kawasaki Masaki Uchida 10.7566/JPSJ.93.023706 http://arxiv.org/abs/2401.00228v1 Parallel-in-time Multilevel Krylov Methods: A Prototype 2023-12-30T13:31:33Z This paper presents a parallel-in-time multilevel iterative method for solving differential algebraic equation, arising from a discretization of linear time-dependent partial differential equation. The core of the method is the multilevel Krylov method, introduced by Erlangga and Nabben~{\it [SIAM J. Sci. Comput., 30(2008), pp. 1572--1595]}. In the method, special time restriction and interpolation operators are proposed to coarsen the time grid and to map functions between fine and coarse time grids. The resulting Galerkin coarse-grid system can be interpreted as time integration of an equivalent differential algebraic equation associated with a larger time step and a modified $θ$-scheme. A perturbed coarse time-grid matrix is used on the coarsest level to decouple the coarsest-level system, allowing full parallelization of the method. Within this framework, spatial coarsening can be included in a natural way, reducing further the size of the coarsest grid problem to solve. Numerical results are presented for the 1- and 2-dimensional heat equation using {\it simulated} parallel implementation, suggesting the potential computational speed-up of up to 9 relative to the single-processor implementation and the speed-up of about 3 compared to the sequential $θ$-scheme. 2023-12-30T13:31:33Z 23 pages Yogi A. Erlangga http://arxiv.org/abs/2401.00246v1 Boosting Large Language Model for Speech Synthesis: An Empirical Study 2023-12-30T14:20:04Z Large language models (LLMs) have made significant advancements in natural language processing and are concurrently extending the language ability to other modalities, such as speech and vision. Nevertheless, most of the previous work focuses on prompting LLMs with perception abilities like auditory comprehension, and the effective approach for augmenting LLMs with speech synthesis capabilities remains ambiguous. In this paper, we conduct a comprehensive empirical exploration of boosting LLMs with the ability to generate speech, by combining pre-trained LLM LLaMA/OPT and text-to-speech synthesis model VALL-E. We compare three integration methods between LLMs and speech synthesis models, including directly fine-tuned LLMs, superposed layers of LLMs and VALL-E, and coupled LLMs and VALL-E using LLMs as a powerful text encoder. Experimental results show that, using LoRA method to fine-tune LLMs directly to boost the speech synthesis capability does not work well, and superposed LLMs and VALL-E can improve the quality of generated speech both in speaker similarity and word error rate (WER). Among these three methods, coupled methods leveraging LLMs as the text encoder can achieve the best performance, making it outperform original speech synthesis models with a consistently better speaker similarity and a significant (10.9%) WER reduction. 2023-12-30T14:20:04Z Hongkun Hao Long Zhou Shujie Liu Jinyu Li Shujie Hu Rui Wang Furu Wei http://arxiv.org/abs/2401.00251v1 The Impact of Thermosolutal Convection on Melting Dynamics of Nano-enhanced Phase Change Materials (NePCM) 2023-12-30T14:36:34Z Nanoparticle-Enhanced Phase Change Materials (NePCM) have been a subject of intensive research owing to their potential for enhanced thermo-physical properties. However, their behavior during phase change processes, such as melting or solidification, remains inadequately understood\@. This investigation focuses on the melting process of NePCM in a square cavity, exploring distinct cases of melting from both the top and bottom sides. The NePCM comprises copper nanoparticles (2 nm in size) suspended in water. Our study involves different combinations of constant temperature boundary conditions and particle volume fractions\@. Utilizing a numerical model based on the one-fluid mixture approach combined with the single-domain enthalpy-porosity model, we account for the phase change process and particles' interaction with the solid-liquid interface. When melting NePCM from the top side, convection effects are suppressed, resulting in a melting process primarily governed by conduction. Both NePCM and pure water melt at the same rate under these conditions. However, melting NePCM from the bottom side induces convection-dominated melting. For pure water, thermal convection leads to the formation of convection cells during melting. Contrastingly, melting NePCM triggers thermosolutal convection due to temperature and particle concentration gradients. The flow cells formed from thermosolutal convection in NePCM differ from those in pure water driven by pure thermal convection. Our simulations reveal that thermosolutal convection contributes to decelerating the solid-liquid interface, thereby prolonging NePCM melting compared to pure water. Surprisingly, the viscosity increase in NePCM plays a minimal role in the deceleration process, contrary to prior literature attributing slow-downs of the melting process of the NePCM primarily to increased viscosity. 2023-12-30T14:36:34Z 61 pages, and 29 Figures Yousef El Hasadi http://arxiv.org/abs/2401.00256v2 Hypergeometric-Type Sequences 2024-04-19T15:18:24Z We introduce hypergeometric-type sequences. They are linear combinations of interlaced hypergeometric sequences (of arbitrary interlacements). We prove that they form a subring of the ring of holonomic sequences. An interesting family of sequences in this class are those defined by trigonometric functions with linear arguments in the index and $π$, such as Chebyshev polynomials, $\left(\sin^2\left(n\,π/4\right)\cdot\cos\left(n\,π/6\right)\right)_n$, and compositions like $\left(\sin\left(\cos(nπ/3)π\right)\right)_n$. We describe an algorithm that computes a hypergeometric-type normal form of a given holonomic $n\text{th}$ term whenever it exists. Our implementation enables us to generate several identities for terms defined via trigonometric functions. 2023-12-30T15:00:54Z 23 pages. To appear in the Journal of Symbolic Computation Bertrand Teguia Tabuguia http://arxiv.org/abs/2401.00259v1 Ultrafast X-ray Diffraction Probe of Coherent Spin-state Dynamics in Molecules 2023-12-30T15:12:15Z We propose an approach to probe coherent spin-state dynamics of molecules using circularly polarized hard x-ray pulses. For the dynamically aligned nitric oxide molecules in a coherent superposition spin-orbit coupled electronic state that can be prepared through stimulated Raman scattering, we demonstrate the capability of ultrafast x-ray diffraction to not only reveal the quantum beating of the coherent spin-state wave packet, but also image the spatial spin density of the molecule. With circularly polarized ultrafast x-ray diffraction signal, we show that the electronic density matrix can be retrieved. The spatio-temporal resolving power of ultrafast x-ray diffraction paves the way for tracking transient spatial wave function in molecular dynamics involving spin degree of freedom. 2023-12-30T15:12:15Z Xiaoyu Mi Ming Zhang Zheng Li http://arxiv.org/abs/2401.00267v2 Hidden Conformal Symmetry from Eight Flavors 2024-01-13T17:11:13Z This proceedings paper extends the scope of our conference talk, where we presented a comprehensive analysis of newly expanded and refined lattice data concerning the SU(3) gauge theory with Nf = 8 light Dirac fermions - a theory positioned near the conformal window boundary. The analysis presented here makes use of a dilaton effective field theory and we delve deeper into the intricacies of the dilaton potential. We aim to clarify the connection between parameters appearing the potential and properties of the underlying gauge theory. 2023-12-30T15:42:02Z 10 pages, 1 figure, Lattice 2023 proceedings. Small changes made to improve clarity of discussion James Ingoldby for the Lattice Strong Dynamics Collaboration http://arxiv.org/abs/2401.00269v1 Sample Robust Scheduling of Electricity-Gas Systems Under Wind Power Uncertainty 2023-12-30T16:06:11Z This paper adopts a two-stage sample robust optimization (SRO) model to address the wind power penetrated unit commitment optimal energy flow (UC-OEF) problem for IEGSs. The two-stage SRO model can be approximately transformed into a computationally efficient form. Specifically, we employ linear decision rules to simplify the proposed UC-OEF model. Moreover, we further enhance the tractability of the simplified model by exploring its structural features and, accordingly, develop a solution method. 2023-12-30T16:06:11Z 10 pages IEEE Trans. Power Syst., vol. 36, no. 6, pp. 5889-5900, Nov. 2021 Rong-Peng Liu Yunhe Hou Yujia Li Shunbo Lei Wei Wei Xiaozhe Wang 10.1109/TPWRS.2021.3081557 http://arxiv.org/abs/2401.00270v2 A number-theoretic problem concerning pseudo-real Riemann surfaces 2024-01-18T21:45:33Z Motivated by their research on automorphism groups of pseudo-real Riemann surfaces, Bujalance, Cirre and Conder have conjectured that there are infinitely many primes $p$ such that $p+2$ has all its prime factors $q\equiv -1$ mod~$(4)$. We use theorems of Landau and Raikov to prove that the number of integers $n\le x$ with only such prime factors $q$ is asymptotic to $cx/\sqrt{\ln x}$ for a specific constant $c=0.4865\ldots$. Heuristic arguments, following Hardy and Littlewood, then yield a conjecture that the number of such primes $p\le x$ is asymptotic to $c'\int_2^x(\ln t)^{-3/2}dt$ for a constant $c'=0.8981\ldots$. The theorem, the conjecture and a similar conjecture applying the Bateman--Horn Conjecture to other pseudo-real Riemann surfaces are supported by evidence from extensive computer searches. 2023-12-30T16:09:00Z 20 pages Gareth A. Jones Alexander K. Zvonkin http://arxiv.org/abs/2401.00285v1 BusReF: Infrared-Visible images registration and fusion focus on reconstructible area using one set of features 2023-12-30T17:32:44Z In a scenario where multi-modal cameras are operating together, the problem of working with non-aligned images cannot be avoided. Yet, existing image fusion algorithms rely heavily on strictly registered input image pairs to produce more precise fusion results, as a way to improve the performance of downstream high-level vision tasks. In order to relax this assumption, one can attempt to register images first. However, the existing methods for registering multiple modalities have limitations, such as complex structures and reliance on significant semantic information. This paper aims to address the problem of image registration and fusion in a single framework, called BusRef. We focus on Infrared-Visible image registration and fusion task (IVRF). In this framework, the input unaligned image pairs will pass through three stages: Coarse registration, Fine registration and Fusion. It will be shown that the unified approach enables more robust IVRF. We also propose a novel training and evaluation strategy, involving the use of masks to reduce the influence of non-reconstructible regions on the loss functions, which greatly improves the accuracy and robustness of the fusion task. Last but not least, a gradient-aware fusion network is designed to preserve the complementary information. The advanced performance of this algorithm is demonstrated by 2023-12-30T17:32:44Z Zeyang Zhang Hui Li Tianyang Xu Xiaojun Wu Josef Kittler http://arxiv.org/abs/2401.00300v1 Exponential and Prescribed-Time Extremum Seeking with Unbiased Convergence 2023-12-30T18:36:27Z We present multivariable extremum seeking (ES) designs that achieve unbiased convergence to the optimum. Two designs are introduced: one with exponential unbiased convergence (unbiased extremum seeker, uES) and the other with user-assignable prescribed-time unbiased convergence (unbiased PT extremum seeker, uPT-ES). In contrast to the conventional ES, which uses persistent sinusoids and results in steady-state oscillations around the optimum, the exponential uES employs an exponentially decaying amplitude in the perturbation signal (for achieving convergence) and an exponentially growing demodulation signal (for making the convergence unbiased). The achievement of unbiased convergence also entails employing an adaptation gain that is sufficiently large in relation to the decay rate of the perturbation amplitude. Stated concisely, the bias is eliminated by having the learning process outpace the waning of the perturbation. The other algorithm, uPT-ES, employs prescribed-time convergent/blow-up functions in place of constant amplitudes of sinusoids, and it also replaces constant-frequency sinusoids with chirp signals whose frequency grows over time. Among the convergence results in the ES literature, uPT-ES may be the strongest yet in terms of the convergence rate (prescribed-time) and accuracy (unbiased). To enhance the robustness of uES to a time-varying optimum, exponential functions are modified to keep oscillations at steady state. Stability analysis of the designs is based on a state transformation, averaging, local exponential/PT stability of the averaged system, local stability of the transformed system, and local exponential/PT stability of the original system. For numerical implementation of the developed ES schemes and comparison with previous ES designs, the problem of source seeking by a two-dimensional velocity-actuated point mass is considered. 2023-12-30T18:36:27Z 16 pages, 7 figures Cemal Tugrul Yilmaz Mamadou Diagne Miroslav Krstic http://arxiv.org/abs/2401.00288v1 Deep Learning for Code Intelligence: Survey, Benchmark and Toolkit 2023-12-30T17:48:37Z Code intelligence leverages machine learning techniques to extract knowledge from extensive code corpora, with the aim of developing intelligent tools to improve the quality and productivity of computer programming. Currently, there is already a thriving research community focusing on code intelligence, with efforts ranging from software engineering, machine learning, data mining, natural language processing, and programming languages. In this paper, we conduct a comprehensive literature review on deep learning for code intelligence, from the aspects of code representation learning, deep learning techniques, and application tasks. We also benchmark several state-of-the-art neural models for code intelligence, and provide an open-source toolkit tailored for the rapid prototyping of deep-learning-based code intelligence models. In particular, we inspect the existing code intelligence models under the basis of code representation learning, and provide a comprehensive overview to enhance comprehension of the present state of code intelligence. Furthermore, we publicly release the source code and data resources to provide the community with a ready-to-use benchmark, which can facilitate the evaluation and comparison of existing and future code intelligence models (https://xcodemind.github.io). At last, we also point out several challenging and promising directions for future research. 2023-12-30T17:48:37Z Yao Wan Yang He Zhangqian Bi Jianguo Zhang Hongyu Zhang Yulei Sui Guandong Xu Hai Jin Philip S. Yu http://arxiv.org/abs/2401.00281v1 Dynamics of oscillator populations with disorder in the coupling phase shifts 2023-12-30T17:00:22Z We study populations of oscillators, all-to-all coupled by means of quenched disordered phase shifts. While there is no traditional synchronization transition with a nonvanishing Kuramoto order parameter, the system demonstrates a specific order as the coupling strength increases. This order is characterized by partial phase locking, which is put into evidence by the introduced correlation order parameter and via frequency entrainment. Simulations with phase oscillators, Stuart-Landau oscillators, and chaotic Roessler oscillators demonstrate similar scaling of the correlation order parameter with the coupling and the system size and also similar behavior of the frequencies with maximal entrainment at some finite coupling. 2023-12-30T17:00:22Z New. J. Physics, v. 26, 023054 (2024) Arkady Pikovsky Franco Bagnoli 10.1088/1367-2630/ad2a80 http://arxiv.org/abs/2401.00235v2 Analytic capacities in Besov spaces 2024-07-13T20:51:32Z We derive new estimates on analytic capacities of finite sequences in the unit disc in Besov spaces with zero smoothness, which sharpen the estimates obtained by N.K.Nikolski in 2005 and, for a range of parameters, are optimal. The work is motivated both from the perspective of complex analysis by the description of sets of zeros/uniqueness, and from the one of matrix analysis/operator theory by estimates on norms of inverses. 2023-12-30T13:51:50Z 21 pages Anton Baranov Michael Hartz Ilgiz Kayumov Rachid Zarouf http://arxiv.org/abs/2401.00261v2 Super-Resolution Microscopy Based on the Inherent Fluctuations of Dye Molecules 2024-08-07T08:45:03Z Fluorescence microscopy is a critical tool across various disciplines, from materials science to biomedical research, yet it is limited by the diffraction limit of resolution. Advanced super-resolution techniques such as localization microscopy and stimulated-emission-depletion microscopy often demand considerable resources. These methods depend heavily on elaborate sample-staining, complex optical systems, or prolonged acquisition periods, and their application in 3D and multicolor imaging presents significant experimental challenges. In the current work, we provide a complete demonstration of a widely accessible super-resolution imaging approach capable of 3D and multicolor imaging. We replace the confocal pinhole with an array of single-photon avalanche diodes and use the microsecond-scale fluctuations of dye molecules as a contrast mechanism. This contrast is transformed into a super-resolved image using a robust and deterministic algorithm. Our technique utilizes natural fluctuations inherent to organic dyes, thereby it does not require engineering of the blinking statistics. Our robust, versatile super-resolution method opens the way to next-generation multimodal imaging and facilitates on-demand super-resolution within a confocal architecture. 2023-12-30T15:26:53Z Alexander Krupinski-Ptaszek Adrian Makowski Aleksandra Mielnicka Monika Pawłowska Ron Tenne Radek Lapkiewicz http://arxiv.org/abs/2401.00263v2 A framework for the valuation of insurance liabilities by production cost 2025-06-01T17:53:55Z This paper sets out a framework for the valuation of insurance liabilities that is intended to be economically realistic, elementary, reasonably practically applicable, and as a special case to provide a basis for the valuation in regulatory solvency systems such as Solvency II and the SST. The valuation framework is based on the cost of producing the liabilities to an insurance company that is subject to solvency regulation (regulatory solvency capital requirements) and insolvency laws (consequences of failure) in finite discrete time. Starting from the replication approach of classical no-arbitrage theory, the framework additionally considers the nature and cost of capital (expressed by a ``financiability condition"), that the liabilities may be required to be fulfilled only ``in sufficiently many cases" (expressed by a ``fulfillment condition"), production using ``fully illiquid" assets in addition to tradables, and the asymmetry between assets and liabilities. We identify necessary and sufficient conditions on the capital investment under which the framework recovers the market prices of tradables, investigate extending production to take account of insolvency, implications of using illiquid assets in the production, and show how Solvency II and SST valuation can be derived with specific assumptions. 2023-12-30T15:30:44Z 35 pages, no figures Christoph Moehr http://arxiv.org/abs/2401.00241v5 Image Super-Resolution Reconstruction Network based on Enhanced Swin Transformer via Alternating Aggregation of Local-Global Features 2025-09-18T06:05:49Z The Swin Transformer image super-resolution (SR) reconstruction network primarily depends on the long-range relationship of the window and shifted window attention to explore features. However, this approach focuses only on global features, ignoring local ones, and considers only spatial interactions, disregarding channel and spatial-channel feature interactions, limiting its nonlinear mapping capability. Therefore, this study proposes an enhanced Swin Transformer network (ESTN) that alternately aggregates local and global features. During local feature aggregation, shift convolution facilitates the interaction between local spatial and channel information. During global feature aggregation, a block sparse global perception module is introduced, wherein spatial information is reorganized and the recombined features are then processed by a dense layer to achieve global perception. Additionally, multiscale self-attention and low-parameter residual channel attention modules are introduced to aggregate information across different scales. Finally, the effectiveness of ESTN on five public datasets and a local attribution map (LAM) are analyzed. Experimental results demonstrate that the proposed ESTN achieves higher average PSNR, surpassing SRCNN, ELAN-light, SwinIR-light, and SMFANER+ models by 2.17dB, 0.13dB, 0.12dB, and 0.1dB, respectively, with LAM further confirming its larger receptive field. ESTN delivers improved quality of SR images. The source code can be found at https://github.com/huangyuming2021/ESTN. 2023-12-30T14:11:08Z Yuming Huang Yingpin Chen Changhui Wu Binhui Song Hui Wang http://arxiv.org/abs/2401.00222v1 Optimization of muonium yield in perforated silica aerogel 2023-12-30T13:10:31Z A muonium consists of a positive muon associated with an orbital electron, and the spontaneous conversion to antimuonium serves as a clear indication of new physics beyond the Standard Model in particle physics.One of the most important aspects in muonium-to-antimuonium conversion experiment (MACE) is to increase the muonium yield in vacuum to challenge the latest limit obtained in 1999. This study focuses on a simulation of the muonium formation and diffusion in the perforated silica aerogel. The independent simulation results can be well validated by experimental data. By optimizing the target geometry, we find a maximum muonium emission efficiency of $7.92(2)\%$ and a maximum vacuum yield of $1.134(2)\%$ with a typical surface muon beam, indicating a 2.6 times and a 2.1 times enhancement, respectively. Our results will pave the way for muonium experiments. 2023-12-30T13:10:31Z 20 pages, 6 figures, 2 tables Shihan Zhao Jian Tang http://arxiv.org/abs/2401.00280v3 Advancing TTP Analysis: Harnessing the Power of Large Language Models with Retrieval Augmented Generation 2024-07-22T02:51:05Z Tactics, Techniques, and Procedures (TTPs) outline the methods attackers use to exploit vulnerabilities. The interpretation of TTPs in the MITRE ATT&CK framework can be challenging for cybersecurity practitioners due to presumed expertise and complex dependencies. Meanwhile, advancements with Large Language Models (LLMs) have led to recent surge in studies exploring its uses in cybersecurity operations. It is, however, unclear how LLMs can be used in an efficient and proper way to provide accurate responses for critical domains such as cybersecurity. This leads us to investigate how to better use two types of LLMs: small-scale encoder-only (e.g., RoBERTa) and larger decoder-only (e.g., GPT-3.5) LLMs to comprehend and summarize TTPs with the intended purposes (i.e., tactics) of a cyberattack procedure. This work studies and compares the uses of supervised fine-tuning (SFT) of encoder-only LLMs vs. Retrieval Augmented Generation (RAG) for decoder-only LLMs (without fine-tuning). Both SFT and RAG techniques presumably enhance the LLMs with relevant contexts for each cyberattack procedure. Our studies show decoder-only LLMs with RAG achieves better performance than encoder-only models with SFT, particularly when directly relevant context is extracted by RAG. The decoder-only results could suffer low `Precision' while achieving high `Recall'. Our findings further highlight a counter-intuitive observation that more generic prompts tend to yield better predictions of cyberattack tactics than those that are more specifically tailored. 2023-12-30T16:56:24Z Reza Fayyazi Rozhina Taghdimi Shanchieh Jay Yang 10.1109/ACSACW65225.2024.00036 http://arxiv.org/abs/2401.00206v2 Functional Inequalities for Brownian Motion on Riemannian Manifolds with Sticky-Reflecting Boundary Diffusion 2024-04-03T08:22:26Z We prove geometric upper bounds for the Poincaré and Logarithmic Sobolev constants for Brownian motion on manifolds with sticky reflecting boundary diffusion i.e. extended Wentzell-type boundary condition under general curvature assumptions on the manifold and its boundary. The method is based on an interpolation involving energy interactions between the boundary and the interior of the manifold. As side results we obtain explicit geometric bounds on the first nontrivial Steklov eigenvalue, for the norm of the boundary trace operator on Sobolev functions, and on the boundary trace logarithmic Sobolev constant. The case of Brownian motion with pure sticky reflection is also treated. 2023-12-30T11:19:23Z Marie Bormann Max von Renesse Feng-Yu Wang http://arxiv.org/abs/2401.00213v2 On the Ground State Quantum Droplet for Large Chemical Potentials 2024-01-16T18:26:44Z In the present work we revisit the problem of the quantum droplet in atomic Bose-Einstein condensates with an eye towards describing its ground state in the large density, so-called Thomas-Fermi limit. We consider the problem as being separable into 3 distinct regions: an inner one, where the Thomas-Fermi approximation is valid, a sharp transition region where the density abruptly drops towards the (vanishing) background value and an outer region which asymptotes to the background value. We analyze the spatial extent of each of these regions, and develop a systematic effective description of the rapid intermediate transition region. Accordingly, we derive a uniformly valid description of the ground state that is found to very accurately match our numerical computations. As an additional application of our considerations, we show that this formulation allows for an analytical approximation of excited states such as the (trapped) dark soliton in the large density limit. 2023-12-30T12:21:52Z 7 pages, 3 figures J. Holmer K. Z. Zhang P. G. Kevrekidis http://arxiv.org/abs/2401.00218v1 Comment on "Multiparty quantum mutual information: An alternative definition" 2023-12-30T13:04:11Z We show that, contrary to the claim by Kumar [Phys. Rev. A 96, 012332 (2017)], the quantum dual total correlation of an $n$-partite quantum state cannot be represented as the quantum relative entropy between $n-1$ copies of the quantum state and the product of $n$ different reduced quantum states for $n \geq 3$. Specifically, we argue that the latter fails to yield a finite value for generalized $n$-partite Greenberger-Horne-Zeilinger states. 2023-12-30T13:04:11Z 2 pages, close to published version Phys. Rev. A 108, 066401 (2023) Jaehak Lee Gibeom Noh Changsuk Noh Jiyong Park 10.1103/PhysRevA.108.066401 http://arxiv.org/abs/2401.00234v1 An experiment to measure electromagnetic memory 2023-12-30T13:51:28Z We describe an experiment to measure the electromagnetic analog of gravitational wave memory, the so-called electromagnetic memory. Whereas gravitational wave memory is a residual displacement of test masses, electromagnetic memory is a residual velocity (i.e. kick) of test charges. The source of gravitational wave memory is energy that is not confined to any bounded spatial region: in the case of binary black hole mergers the emitted energy of gravitational radiation as well as the recoil energy of the final black hole. Similarly, electromagnetic memory requires a source whose charges are not confined to any bounded spatial region. While particle beams can provide unbounded charges, their currents are too small to be practical for such an experiment. Instead we propose a short microwave pulse applied to the center of a long dipole antenna. In this way the measurement of the kick can be done quickly enough that the finite size of the antenna does not come into play and it acts for our purposes the same as if it were an infinite antenna. 2023-12-30T13:51:28Z Lydia Bieri David Garfinkle http://arxiv.org/abs/2401.00238v1 How to Evaluate Coreference in Literary Texts? 2023-12-30T14:02:36Z In this short paper, we examine the main metrics used to evaluate textual coreference and we detail some of their limitations. We show that a unique score cannot represent the full complexity of the problem at stake, and is thus uninformative, or even misleading. We propose a new way of evaluating coreference, taking into account the context (in our case, the analysis of fictions, esp. novels). More specifically, we propose to distinguish long coreference chains (corresponding to main characters), from short ones (corresponding to secondary characters), and singletons (isolated elements). This way, we hope to get more interpretable and thus more informative results through evaluation. 2023-12-30T14:02:36Z Presented as a poster at the conference CHR2023 (non archival) Ana-Isabel Duron-Tejedor Pascal Amsili Thierry Poibeau http://arxiv.org/abs/2401.00239v1 Strain induced electronic and magnetic transition in S = 3/2 antiferromagnetic spin chain compound LaCrS3 2023-12-30T14:03:46Z Exploring the physics of low-dimensional spin systems and their pressure-driven electronic and magnetic transitions are thriving research field in modern condensed matter physics. In this context, recently antiferromagnetic Cr-based compounds such as CrI3, CrBr3, CrGeTe3 have been investigated experimentally and theoretically for their possible spintronics applications. Motivated by the fundamental and industrial importance of these materials, we theoretically studied the electronic and magnetic properties of a relatively less explored Cr-based chalcogenide, namely LaCrS3 where 2D layers of magnetic Cr3+ ions form a rectangular lattice. We employed density functional theory + Hubbard U approach in conjunction with constrained random-phase approximation (cRPA) where the later was used to estimate the strength of U. Our findings at ambient pressure show that the system exhibits semiconducting antiferromagnetic ground state with a gap of 0.5 eV and large Cr moments that corresponds to nominal S=3/2 spin-state. The 1st nearest neighbor (NN) interatomic exchange coupling (J1) is found to be strongly antiferromagnetic (AFM), while 2nd NN couplings are relatively weaker ferromagnetic (FM), making this system a candidate for 1D non-frustrated antiferromagnetic spin-chain family of materials. Based on orbital resolved interactions, we demonstrated the reason behind two different types of interactions among 1st and 2nd NN despite their very similar bond lengths. We observe a significant spin-orbit coupling effect, giving rise to a finite magneto crystalline anisotropy, and Dzyaloshinskii-Moriya (DM) interaction. Further, we found that by applying uniaxial tensile strain along crystallographic a and b-axis, LaCrS3 exhibits a magnetic transition to a semi-conducting FM ground state, while compression gives rise to the realization of novel gapless semiconducting antiferromagnetic ground state. 2023-12-30T14:03:46Z Accepted in Phys. Rev. B Kuldeep Kargeti Aadit Sen S. K. Panda http://arxiv.org/abs/2401.00243v1 Uncertainty-Penalized Reinforcement Learning from Human Feedback with Diverse Reward LoRA Ensembles 2023-12-30T14:14:14Z Reinforcement learning from human feedback (RLHF) emerges as a promising paradigm for aligning large language models (LLMs). However, a notable challenge in RLHF is overoptimization, where beyond a certain threshold, the pursuit of higher rewards leads to a decline in human preferences. In this paper, we observe the weakness of KL regularization which is commonly employed in existing RLHF methods to address overoptimization. To mitigate this limitation, we scrutinize the RLHF objective in the offline dataset and propose uncertainty-penalized RLHF (UP-RLHF), which incorporates uncertainty regularization during RL-finetuning. To enhance the uncertainty quantification abilities for reward models, we first propose a diverse low-rank adaptation (LoRA) ensemble by maximizing the nuclear norm of LoRA matrix concatenations. Then we optimize policy models utilizing penalized rewards, determined by both rewards and uncertainties provided by the diverse reward LoRA ensembles. Our experimental results, based on two real human preference datasets, showcase the effectiveness of diverse reward LoRA ensembles in quantifying reward uncertainty. Additionally, uncertainty regularization in UP-RLHF proves to be pivotal in mitigating overoptimization, thereby contributing to the overall performance. 2023-12-30T14:14:14Z 10 pages, 5 figures, Yuanzhao Zhai Han Zhang Yu Lei Yue Yu Kele Xu Dawei Feng Bo Ding Huaimin Wang http://arxiv.org/abs/2401.00266v3 Two-cardinal derived topologies, indescribability and Ramseyness 2024-02-13T14:00:07Z We introduce a natural two-cardinal version of Bagaria's sequence of derived topologies on ordinals. We prove that for our sequence of two-cardinal derived topologies, limit points of sets can be characterized in terms of a new iterated form of pairwise simultaneous reflection of certain kinds of stationary sets, the first few instances of which are often equivalent to notions related to strong stationarity, which has been studied previously in the context of strongly normal ideals. The non-discreteness of these two-cardinal derived topologies can be obtained from certain two-cardinal indescribability hypotheses, which follow from local instances of supercompactness. Additionally, we answer several questions posed by the first author, Peter Holy and Philip White on the relationship between Ramseyness and indescribability in both the cardinal context and in the two-cardinal context. 2023-12-30T15:37:16Z Added citation to the work of Catalina Torres Brent Cody Chris Lambie-Hanson Jing Zhang http://arxiv.org/abs/2401.00268v1 COMMA: Co-Articulated Multi-Modal Learning 2023-12-30T15:47:36Z Pretrained large-scale vision-language models such as CLIP have demonstrated excellent generalizability over a series of downstream tasks. However, they are sensitive to the variation of input text prompts and need a selection of prompt templates to achieve satisfactory performance. Recently, various methods have been proposed to dynamically learn the prompts as the textual inputs to avoid the requirements of laboring hand-crafted prompt engineering in the fine-tuning process. We notice that these methods are suboptimal in two aspects. First, the prompts of the vision and language branches in these methods are usually separated or uni-directionally correlated. Thus, the prompts of both branches are not fully correlated and may not provide enough guidance to align the representations of both branches. Second, it's observed that most previous methods usually achieve better performance on seen classes but cause performance degeneration on unseen classes compared to CLIP. This is because the essential generic knowledge learned in the pretraining stage is partly forgotten in the fine-tuning process. In this paper, we propose Co-Articulated Multi-Modal Learning (COMMA) to handle the above limitations. Especially, our method considers prompts from both branches to generate the prompts to enhance the representation alignment of both branches. Besides, to alleviate forgetting about the essential knowledge, we minimize the feature discrepancy between the learned prompts and the embeddings of hand-crafted prompts in the pre-trained CLIP in the late transformer layers. We evaluate our method across three representative tasks of generalization to novel classes, new target datasets and unseen domain shifts. Experimental results demonstrate the superiority of our method by exhibiting a favorable performance boost upon all tasks with high efficiency. 2023-12-30T15:47:36Z Accepted to AAAI2024. Code is available at https://github.com/hulianyuyy/COMMA Lianyu Hu Liqing Gao Zekang Liu Chi-Man Pun Wei Feng http://arxiv.org/abs/2401.00277v1 Ultracold Neutrons in the Low Curvature Limit: Remarks on the post-Newtonian effects 2023-12-30T16:45:56Z Ultracold neutrons are great experimental tools to explore the gravitational interaction in the regime of quantized states. From a theoretical perspective, starting from a Dirac equation in curved spacetime, we applied a perturbative scheme to systematically derive the non-relativistic Schrödinger equation that governs the evolution of the neutron's wave function in the Earth's gravitational field. At the lowest order, this procedure reproduces a Schrödinger system affected by a linear Newtonian potential, but corrections due to both curvature and relativistic effects are present. Here, we argue that one should be very careful when going one step further in the perturbative expansion. Proceeding methodically with the help of the Foldy-Wouthuysen transformation and a formal post-Newtonian $c^{-2}-$expansion, we derive the non-relativistic Hamiltonian for a generic static spacetime. By employing Fermi coordinates within this framework, we calculate the next-to-leading order corrections to the neutron's energy spectrum. Finally, we evaluate them for typical experimental configurations, such as that of qBOUNCE, and note that, while the current precision for observations of ultracold neutrons may not yet enable to probe them, they could still be relevant in the future or in alternative circumstances. 2023-12-30T16:45:56Z Benjamin Koch Enrique Muñoz Alessandro Santoni http://arxiv.org/abs/2401.00282v1 Deep Generative Symbolic Regression 2023-12-30T17:05:31Z Symbolic regression (SR) aims to discover concise closed-form mathematical equations from data, a task fundamental to scientific discovery. However, the problem is highly challenging because closed-form equations lie in a complex combinatorial search space. Existing methods, ranging from heuristic search to reinforcement learning, fail to scale with the number of input variables. We make the observation that closed-form equations often have structural characteristics and invariances (e.g., the commutative law) that could be further exploited to build more effective symbolic regression solutions. Motivated by this observation, our key contribution is to leverage pre-trained deep generative models to capture the intrinsic regularities of equations, thereby providing a solid foundation for subsequent optimization steps. We show that our novel formalism unifies several prominent approaches of symbolic regression and offers a new perspective to justify and improve on the previous ad hoc designs, such as the usage of cross-entropy loss during pre-training. Specifically, we propose an instantiation of our framework, Deep Generative Symbolic Regression (DGSR). In our experiments, we show that DGSR achieves a higher recovery rate of true equations in the setting of a larger number of input variables, and it is more computationally efficient at inference time than state-of-the-art RL symbolic regression solutions. 2023-12-30T17:05:31Z In the proceedings of the Eleventh International Conference on Learning Representations (ICLR 2023). https://iclr.cc/virtual/2023/poster/11782 International Conference on Learning Representations (ICLR), 2023 Samuel Holt Zhaozhi Qian Mihaela van der Schaar http://arxiv.org/abs/2401.00284v1 Evaluation is all you need. Prompting Generative Large Language Models for Annotation Tasks in the Social Sciences. A Primer using Open Models 2023-12-30T17:22:01Z This paper explores the use of open generative Large Language Models (LLMs) for annotation tasks in the social sciences. The study highlights the challenges associated with proprietary models, such as limited reproducibility and privacy concerns, and advocates for the adoption of open (source) models that can be operated on independent devices. Two examples of annotation tasks, sentiment analysis in tweets and identification of leisure activities in childhood aspirational essays are provided. The study evaluates the performance of different prompting strategies and models (neural-chat-7b-v3-2, Starling-LM-7B-alpha, openchat_3.5, zephyr-7b-alpha and zephyr-7b-beta). The results indicate the need for careful validation and tailored prompt engineering. The study highlights the advantages of open models for data privacy and reproducibility. 2023-12-30T17:22:01Z Maximilian Weber Merle Reichardt http://arxiv.org/abs/2401.00289v1 ASL Champ!: A Virtual Reality Game with Deep-Learning Driven Sign Recognition 2023-12-30T17:55:30Z We developed an American Sign Language (ASL) learning platform in a Virtual Reality (VR) environment to facilitate immersive interaction and real-time feedback for ASL learners. We describe the first game to use an interactive teaching style in which users learn from a fluent signing avatar and the first implementation of ASL sign recognition using deep learning within the VR environment. Advanced motion-capture technology powers an expressive ASL teaching avatar within an immersive three-dimensional environment. The teacher demonstrates an ASL sign for an object, prompting the user to copy the sign. Upon the user's signing, a third-party plugin executes the sign recognition process alongside a deep learning model. Depending on the accuracy of a user's sign production, the avatar repeats the sign or introduces a new one. We gathered a 3D VR ASL dataset from fifteen diverse participants to power the sign recognition model. The proposed deep learning model's training, validation, and test accuracy are 90.12%, 89.37%, and 86.66%, respectively. The functional prototype can teach sign language vocabulary and be successfully adapted as an interactive ASL learning platform in VR. 2023-12-30T17:55:30Z 36 pages, 9 figures Md Shahinur Alam Jason Lamberton Jianye Wang Carly Leannah Sarah Miller Joseph Palagano Myles de Bastion Heather L. Smith Melissa Malzkuhn Lorna C. Quandt http://arxiv.org/abs/2401.00297v1 A Novel Reinforcement Learning Routing Algorithm for Congestion Control in Complex Networks 2023-12-30T18:21:13Z Despite technological advancements, the significance of interdisciplinary subjects like complex networks has grown. Exploring communication within these networks is crucial, with traffic becoming a key concern due to the expanding population and increased need for connections. Congestion tends to originate in specific network areas but quickly proliferates throughout. Consequently, understanding the transition from a flow-free state to a congested state is vital. Numerous studies have delved into comprehending the emergence and control of congestion in complex networks, falling into three general categories: soft strategies, hard strategies, and resource allocation strategies. This article introduces a routing algorithm leveraging reinforcement learning to address two primary objectives: congestion control and optimizing path length based on the shortest path algorithm, ultimately enhancing network throughput compared to previous methods. Notably, the proposed method proves effective not only in Barabási-Albert scale-free networks but also in other network models such as Watts-Strogatz (small-world) and Erdös-Rényi (random network). Simulation experiment results demonstrate that, across various traffic scenarios and network topologies, the proposed method can enhance efficiency criteria by up to 30% while reducing maximum node congestion by five times. 2023-12-30T18:21:13Z 15 pages, 8 figures, under review at Journal of Systems Science & Complexity Seyed Hassan Yajadda Farshad Safaei http://arxiv.org/abs/2401.00298v1 Principal-Agent Reward Shaping in MDPs 2023-12-30T18:30:44Z Principal-agent problems arise when one party acts on behalf of another, leading to conflicts of interest. The economic literature has extensively studied principal-agent problems, and recent work has extended this to more complex scenarios such as Markov Decision Processes (MDPs). In this paper, we further explore this line of research by investigating how reward shaping under budget constraints can improve the principal's utility. We study a two-player Stackelberg game where the principal and the agent have different reward functions, and the agent chooses an MDP policy for both players. The principal offers an additional reward to the agent, and the agent picks their policy selfishly to maximize their reward, which is the sum of the original and the offered reward. Our results establish the NP-hardness of the problem and offer polynomial approximation algorithms for two classes of instances: Stochastic trees and deterministic decision processes with a finite horizon. 2023-12-30T18:30:44Z Full version of a paper accepted to AAAI'24 Omer Ben-Porat Yishay Mansour Michal Moshkovitz Boaz Taitler http://arxiv.org/abs/2401.00231v1 Survivability of Amorphous Ice in Comets Depends on the Latent Heat of Crystallization of Impure Water Ice 2023-12-30T13:44:23Z Comets would have amorphous ice rather than crystalline one at the epoch of their accretion. Cometary ice contains some impurities that govern the latent heat of ice crystallization, $L_{\rm cry}$. However, it is still controversial whether the crystallization process is exothermic or endothermic. In this study, we perform one-dimensional simulations of the thermal evolution of km-sized comets and investigate the effect of the latent heat. We find that the depth where amorphous ice can survive significantly depends on the latent heat of ice crystallization. Assuming the cometary radius of 2 km, the depth of the amorphous ice mantle is approximately 100 m when the latent heat is positive (i.e., the exothermic case with $L_{\rm cry} = + 9 \times 10^{4}$ J/kg). In contrast, when we consider the impure ice representing the endothermic case with $L_{\rm cry} = - 9 \times 10^{4}$ J/kg, the depth of the amorphous ice mantle could exceed 1 km. Although our numerical results indicate that these depths depend on the size and the accretion age of comets, the depth in a comet with the negative latent heat is a few to several times larger than the positive case for a given comet size. This work suggests that the spatial distribution of the ice crystallinity in a comet nucleus depends on the latent heat, which can be different from the previous estimates assuming pure water ice. 2023-12-30T13:44:23Z 15 pages, 10 figures. Accepted for publication in PASJ Sota Arakawa Shigeru Wakita 10.1093/pasj/psad086 http://arxiv.org/abs/2401.00207v2 A unified structure-preserving parametric finite element method for anisotropic surface diffusion 2024-08-31T13:14:52Z We propose and analyze a unified structure-preserving parametric finite element method (SP-PFEM) for the anisotropic surface diffusion of curves in two dimensions $(d=2)$ and surfaces in three dimensions $(d=3)$ with an arbitrary anisotropic surface energy density $γ(\boldsymbol{n})$, where $\boldsymbol{n}\in \mathbb{S}^{d-1}$ represents the outward unit vector. By introducing a novel unified surface energy matrix $\boldsymbol{G}_k(\boldsymbol{n})$ depending on $γ(\boldsymbol{n})$, the Cahn--Hoffman $\boldsymbolξ$-vector and a stabilizing function $k(\boldsymbol{n}):\ \mathbb{S}^{d-1}\to {\mathbb R}$, we obtain a unified and conservative variational formulation for the anisotropic surface diffusion via different surface differential operators including the surface gradient operator, the surface divergence operator and the surface Laplace--Beltrami operator. A SP-PFEM discretization is presented for the variational problem. In order to establish the unconditional energy stability of the proposed SP-PFEM under a very mild condition on $γ(\boldsymbol{n})$, we propose a new framework via {\sl local energy estimate} for proving energy stability/structure-preserving properties of the parametric finite element method for the anisotropic surface diffusion. This framework sheds light on how to prove unconditional energy stability of other numerical methods for geometric partial differential equations. Extensive numerical results are reported to demonstrate the efficiency and accuracy as well as structure-preserving properties of the proposed SP-PFEM for the anisotropic surface diffusion with arbitrary anisotropic surface energy density $γ(\boldsymbol{n})$ arising from different applications. 2023-12-30T11:26:32Z Weizhu Bao Yifei Li http://arxiv.org/abs/2401.00202v2 Roots of identity in finite groups of Lie type 2024-05-28T06:15:12Z Given an integer $M\geq 2$, we deploy the generating function techniques to compute the number of $M$-th roots of identity in some of the well-known finite groups of Lie type, more precisely for finite general linear groups, symplectic groups, orthogonal groups of all types and unitary groups over finite fields of odd characteristics. 2023-12-30T10:57:57Z Revised version; Suggestions are welcome; Saikat Panja http://arxiv.org/abs/2401.00258v2 Phase diagram and critical behavior of Hubbard model on the square-hexagon-octagon lattice 2024-06-16T15:18:34Z Employing the projective formalism of determinant quantum Monte Carlo (DQMC) simulations, we meticulously explore the ground-state phase diagram and critical behavior of the half-filled Hubbard model on a square-hexagon-octagon (SHO) lattice. This lattice, a two-dimensional (2D) structure comprising squares, hexagons, and octagons, is representative of the biphenylene network (BPN). Our findings reveal an intriguing ground-state phase diagram, featuring an antiferromagnetic (AFM) Mott insulating phase enveloped by three valence-bond solid-like (VBS-like) insulating phases. Analyzing the single-particle gap, spin gap, and single-particle spectral function, we observe that the metallic state in the noninteracting case becomes unstable under the influence of Hubbard U. This interaction drives the system into a hexagon insulating phase before transitioning into an AFM Mott insulating phase. To quantify the critical exponents, we use finite-size scaling techniques. The critical exponents of quantum critical points between the AFM Mott insulating phase and two insulating phases, plaquette insulator and ethylene insulator, closely align with the 3D O(3) universality class. However, the critical exponents of quantum critical points between the hexagon insulating phase and the AFM Mott insulating phase deviate from the 3D O(3) universality class. This deviation is a finite-size effect and can be attributed to the coupling between the fluctuations of magnetic order parameter and very low-energy fermionic excitations. Our comprehensive study not only advances the understanding of correlation effects on the SHO lattice but also sheds light on the less-explored critical exponents in weakly insulating quantum critical point. 2023-12-30T15:11:07Z Phys. Rev. B 109, 155122 (2024) Xinwei Jia Dao-Xin Yao Han-Qing Wu 10.1103/PhysRevB.109.155122 http://arxiv.org/abs/2401.00299v3 Partitioning the hypercube into smaller hypercubes 2024-11-07T05:30:36Z Denote by Q_d the d-dimensional hypercube. Addressing a recent question we estimate the number of ways the vertex set of Q_d can be partitioned into vertex disjoint smaller cubes. Among other results, we prove that the asymptotic order of this function is not much larger than the number of perfect matchings of Q_d. We also describe several new (and old) questions. 2023-12-30T18:35:55Z Proofs slightly shortened and referee comments addressed Illinois J. Math. 69 (1), 109-122, (2025) Noga Alon Jozsef Balogh Vladimir N. Potapov 10.1215/00192082-11792788 http://arxiv.org/abs/2401.00244v1 Non-smoothable $\mathbb{Z}/p$-actions on nuclei 2023-12-30T14:17:59Z In this article we construct examples of non-smoothable $\mathbb{Z}/p$-actions on indefinite spin 4-manifolds with boundary for all primes $p\geq 5$. For example, we show that for each prime $p\geq 5$ and each $n\geq 1$ there exists a locally linear $\mathbb{Z}/p$-action on the Gompf nucleus $N(2pn)$ which is not smoothable with respect to any smooth structure on $N(2pn)$. Furthermore we investigate the behavior of these actions under two different types of equivariant stabilizations with $S^{2}\times S^{2}$, namely \emph{free} and \emph{homologically trivial} stabilizations -- in particular we show that our non-smoothable $\mathbb{Z}/p$-action on $N(2pn)$ remains non-smoothable after $2n-2$ free stabilizations, and after arbitrarily many homologically trivial stabilizations. We also show that free stabilizations satisfy a Wall stabilization principle in the sense that any non-smoothable $\mathbb{Z}/p$-action becomes smoothable after some finite number free stabilizations (under certain assumptions), whereas our aforementioned result implies that homologically trivial stabilizations do not satisfy this property. The proofs of these results use equivariant $κ$-invariants defined by the author in \cite{Mon22}, calculations of equivariant $η$-invariants for the odd signature and Dirac operators on Seifert-fibered spaces, as well as an analysis of the geometric $S^{1}$-action on the Seiberg-Witten moduli spaces of Seifert-fibered spaces induced by rotation in the fibers, which may be of independent interest. 2023-12-30T14:17:59Z 42 pages, 2 figures. Comments welcome! Imogen Montague http://arxiv.org/abs/2401.00249v2 Forecasting CPI inflation under economic policy and geopolitical uncertainties 2024-07-02T14:46:18Z Forecasting consumer price index (CPI) inflation is of paramount importance for both academics and policymakers at the central banks. This study introduces a filtered ensemble wavelet neural network (FEWNet) to forecast CPI inflation, which is tested on BRIC countries. FEWNet breaks down inflation data into high and low-frequency components using wavelets and utilizes them along with other economic factors (economic policy uncertainty and geopolitical risk) to produce forecasts. All the wavelet-transformed series and filtered exogenous variables are fed into downstream autoregressive neural networks to make the final ensemble forecast. Theoretically, we show that FEWNet reduces the empirical risk compared to fully connected autoregressive neural networks. FEWNet is more accurate than other forecasting methods and can also estimate the uncertainty in its predictions due to its capacity to effectively capture non-linearities and long-range dependencies in the data through its adaptable architecture. This makes FEWNet a valuable tool for central banks to manage inflation. 2023-12-30T14:34:22Z International Journal of Forecasting, 2024 Shovon Sengupta Tanujit Chakraborty Sunny Kumar Singh 10.1016/j.ijforecast.2024.08.005 http://arxiv.org/abs/2401.00248v4 Promoting Segment Anything Model towards Highly Accurate Dichotomous Image Segmentation 2025-03-25T12:24:08Z The Segment Anything Model (SAM) represents a significant breakthrough into foundation models for computer vision, providing a large-scale image segmentation model. However, despite SAM's zero-shot performance, its segmentation masks lack fine-grained details, particularly in accurately delineating object boundaries. Therefore, it is both interesting and valuable to explore whether SAM can be improved towards highly accurate object segmentation, which is known as the dichotomous image segmentation (DIS) task. To address this issue, we propose DIS-SAM, which advances SAM towards DIS with extremely accurate details. DIS-SAM is a framework specifically tailored for highly accurate segmentation, maintaining SAM's promptable design. DIS-SAM employs a two-stage approach, integrating SAM with a modified advanced network that was previously designed to handle the prompt-free DIS task. To better train DIS-SAM, we employ a ground truth enrichment strategy by modifying original mask annotations. Despite its simplicity, DIS-SAM significantly advances the SAM, HQ-SAM, and Pi-SAM ~by 8.5%, ~6.9%, and ~3.7% maximum F-measure. Our code at https://github.com/Tennine2077/DIS-SAM 2023-12-30T14:24:33Z Xianjie Liu Keren Fu Yao Jiang Qijun Zhao http://arxiv.org/abs/2401.00211v2 Open-TI: Open Traffic Intelligence with Augmented Language Model 2024-12-04T20:18:30Z Transportation has greatly benefited the cities' development in the modern civilization process. Intelligent transportation, leveraging advanced computer algorithms, could further increase people's daily commuting efficiency. However, intelligent transportation, as a cross-discipline, often requires practitioners to comprehend complicated algorithms and obscure neural networks, bringing a challenge for the advanced techniques to be trusted and deployed in practical industries. Recognizing the expressiveness of the pre-trained large language models, especially the potential of being augmented with abilities to understand and execute intricate commands, we introduce Open-TI. Serving as a bridge to mitigate the industry-academic gap, Open-TI is an innovative model targeting the goal of Turing Indistinguishable Traffic Intelligence, it is augmented with the capability to harness external traffic analysis packages based on existing conversations. Marking its distinction, Open-TI is the first method capable of conducting exhaustive traffic analysis from scratch - spanning from map data acquisition to the eventual execution in complex simulations. Besides, Open-TI is able to conduct task-specific embodiment like training and adapting the traffic signal control policies (TSC), explore demand optimizations, etc. Furthermore, we explored the viability of LLMs directly serving as control agents, by understanding the expected intentions from Open-TI, we designed an agent-to-agent communication mode to support Open-TI conveying messages to ChatZero (control agent), and then the control agent would choose from the action space to proceed the execution. We eventually provide the formal implementation structure, and the open-ended design invites further community-driven enhancements. 2023-12-30T11:50:11Z Published on International Journal of Machine Learning and Cybernetics, Preview version: https://rdcu.be/dHu0b, Github: https://github.com/DaRL-LibSignal/OpenTI Longchao Da Kuanru Liou Tiejin Chen Xuesong Zhou Xiangyong Luo Yezhou Yang Hua Wei 10.1007/s13042-024-02190-8 http://arxiv.org/abs/2401.00275v3 An $\ell^1$-Plug-and-Play Approach for MPI Using a Zero Shot Denoiser with Evaluation on the 3D Open MPI Dataset 2025-04-28T09:59:06Z Objective: Magnetic particle imaging (MPI) is an emerging medical imaging modality which has gained increasing interest in recent years. Among the benefits of MPI are its high temporal resolution, and that the technique does not expose the specimen to any kind of ionizing radiation. It is based on the non-linear response of magnetic nanoparticles to an applied magnetic field. From the electric signal measured in receive coils, the particle concentration has to be reconstructed. Due to the ill-posedness of the reconstruction problem, various regularization methods have been proposed for reconstruction ranging from early stopping methods, via classical Tikhonov regularization and iterative methods to modern machine learning approaches. In this work, we contribute to the latter class: we propose a plug-and-play approach based on a generic zero-shot denoiser with an $\ell^1$-prior. Approach: We validate the reconstruction parameters of the method on a hybrid dataset and compare it with the baseline Tikhonov, DIP and the previous PP-MPI, which is a plug-and-play method with denoiser trained on MPI-friendly data. Main results: We offer a quantitative and qualitative evaluation of the zero-shot plug-and-play approach on the 3D Open MPI dataset. Moreover, we show the quality of the approach with different levels of preprocessing of the data. Significance: The proposed method employs a zero-shot denoiser which has not been trained for the MPI task and therefore saves the cost for training. Moreover, it offers a method that can be potentially applied in future MPI contexts. 2023-12-30T16:27:43Z 24 pages, 7 figures, additional supplementary material (78 pages total) Phys. Med. Biol. (70) 025028 (2025) Vladyslav Gapyak Corinna Rentschler Thomas März Andreas Weinmann 10.1088/1361-6560/ada5a1