{"id": "arxiv:2401.00001v1", "text": "This study presents an analytical approach to sector rotation, leveraging both factor models and fundamental metrics. We initiate with a systematic classification of sectors, followed by an empirical investigation into their returns. Through factor analysis, the paper underscores the significance of momentum and short-term reversion in dictating sectoral shifts. A subsequent in-depth fundamental analysis evaluates metrics such as PE, PB, EV-to-EBITDA, Dividend Yield, among others. Our primary contribution lies in developing a predictive framework based on these fundamental indicators. The constructed models, post rigorous training, exhibit noteworthy predictive capabilities. The findings furnish a nuanced understanding of sector rotation strategies, with implications for asset management and portfolio construction in the financial domain."} {"id": "arxiv:2401.00006v3", "text": "Building embodied agents on integrating Large Language Models (LLMs) and Reinforcement Learning (RL) have revolutionized human-AI interaction: researchers can now leverage language instructions to plan decision-making for open-ended tasks. However, existing research faces challenges in meeting the requirement of open-endedness. They typically either train LLM/RL models to adapt to a fixed counterpart, limiting exploration of novel skills and hindering the efficacy of human-AI interaction. To this end, we present OpenPAL, a co-training framework comprising two stages: (1) fine-tuning a pre-trained LLM to translate human instructions into goals for planning, and goal-conditioned training a policy for decision-making; (2) co-training to align the LLM and policy, achieving instruction open-endedness. We conducted experiments using Contra, an open-ended FPS game, demonstrating that an agent trained with OpenPAL not only comprehends arbitrary instructions but also exhibits efficient execution. These results suggest that OpenPAL holds the potential to construct open-ended embodied agents in practical scenarios."} {"id": "arxiv:2401.00011v1", "text": "Recent years have seen a lot of progress in algorithms for learning parameters of spreading dynamics from both full and partial data. Some of the remaining challenges include model selection under the scenarios of unknown network structure, noisy data, missing observations in time, as well as an efficient incorporation of prior information to minimize the number of samples required for an accurate learning. Here, we introduce a universal learning method based on scalable dynamic message-passing technique that addresses these challenges often encountered in real data. The algorithm leverages available prior knowledge on the model and on the data, and reconstructs both network structure and parameters of a spreading model. We show that a linear computational complexity of the method with the key model parameters makes the algorithm scalable to large network instances."} {"id": "arxiv:2401.00017v1", "text": "I use QAOA to solve the Hamiltonian Circle problem. First, inspired by Lucas, I define the QUBO form of Hamiltonian Cycle and transform it to a quantum circuit by embedding the problem of $n$ vertices to an encoding of $(n-1)^2$ qubits. Then, I calcluate the spectrum of the cost hamiltonian for both triangle case and square case and justify my definition. I also write a python program to generate the cost hamiltonian automatically for finding the hamiltonian cycle in an arbitrary graph. I test the correctess of the hamailtonian by analyze their energy spectrums. Since the $(n-1)^2$ embedding limit my simulation of graph size to be less than $5$, I decide to test the correctness, only for small and simple graph in this project. I implement the QAOA algorithm using qiskit and run the simulation for the triangle case and the square case, which are easy to test the correctness, both with and without noise. A very interesting result I got is that for the square case, the QAOA get much better result on a noisy simulator than a noiseless simulator. The explanation for this phenomena require further investigation, perhaps quantum noise can actually be helpful, rather than harmful in the annealing algorithms. I also use two different kinds of mixer, $R_x$ mixer and $R_y$ circuit to run the simulation. It turns out that $R_x$ mixer performs much better than $R_y$ mixer in this problem."} {"id": "arxiv:2401.00026v1", "text": "We reaffirm the claim of Lee et al. [preceding Comment, Phys. Rev. A 108, 066401 (2023)] that the expression of quantum dual total correlation of a multipartite system in terms of quantum relative entropy as proposed in previous work [A. Kumar, Phys. Rev. A 96, 012332 (2017)] is not correct. We provide alternate expression(s) of quantum dual total correlation in terms of quantum relative entropy. We, however, prescribe that in computing quantum dual total correlation one should use its expression in terms of von Neumann entropy."} {"id": "arxiv:2401.00029v3", "text": "Estimating the 6D object pose from a single RGB image often involves noise and indeterminacy due to challenges such as occlusions and cluttered backgrounds. Meanwhile, diffusion models have shown appealing performance in generating high-quality images from random noise with high indeterminacy through step-by-step denoising. Inspired by their denoising capability, we propose a novel diffusion-based framework (6D-Diff) to handle the noise and indeterminacy in object pose estimation for better performance. In our framework, to establish accurate 2D-3D correspondence, we formulate 2D keypoints detection as a reverse diffusion (denoising) process. To facilitate such a denoising process, we design a Mixture-of-Cauchy-based forward diffusion process and condition the reverse process on the object features. Extensive experiments on the LM-O and YCB-V datasets demonstrate the effectiveness of our framework."} {"id": "arxiv:2401.00046v3", "text": "Based on the covariant underdamped and overdamped Langevin equations with Stratonovich coupling to multiplicative noises and the associated Fokker-Planck equations on Riemannian manifold, we present the first law of stochastic thermodynamics on the trajectory level. The corresponding fluctuation theorems are also established, with the total entropy production of the Brownian particle and the heat reservoir playing the role of dissipation function."} {"id": "arxiv:2401.00065v1", "text": "Addressing the uncertainty and variability in the quality of 3D printed metals can further the wide spread use of this technology. Process mapping for new alloys is crucial for determining optimal process parameters that consistently produce acceptable printing quality. Process mapping is typically performed by conventional methods and is used for the design of experiments and ex situ characterization of printed parts. On the other hand, in situ approaches are limited because their observable features are limited and they require complex high-cost setups to obtain temperature measurements to boost accuracy. Our method relaxes these limitations by incorporating the temporal features of molten metal dynamics during laser-metal interactions using video vision transformers and high-speed imaging. Our approach can be used in existing commercial machines and can provide in situ process maps for efficient defect and variability quantification. The generalizability of the approach is demonstrated by performing cross-dataset evaluations on alloys with different compositions and intrinsic thermofluid properties."} {"id": "arxiv:2401.00066v1", "text": "For $X$ a smooth projective variety, the quantum cohomology ring $QH^*(X)$ is a deformation of the usual cohomology ring $H^*(X)$, where the product structure is modified to incorporate quantum corrections. These correction terms are defined using Gromov-Witten invariants. When $X$ is toric with the geometric quotient description $V /\\!/ T$, the cohomology ring $H^*(V /\\!/T)$ also has the structure of a quantum $H^*(T)$-module. In this paper, we give a new deformation using quasimap invariants with a light point. This defines $H^*(T)$-module structure on $H^*(X)$ through a modified version of the WDVV equations. Using the Atiyah-Bott localization theorem, we explicitly compute this structure for the Hirzebruch surface of type 2. We conjecture that this new quantum module structure is isomorphic to the natural module structure of the Batyrev ring for a semipositive toric variety."} {"id": "arxiv:2401.00076v1", "text": "Seasonal influenza causes on average 425,000 hospitalizations and 32,000 deaths per year in the United States. Forecasts of influenza-like illness (ILI) -- a surrogate for the proportion of patients infected with influenza -- support public health decision making. The goal of an ensemble forecast of ILI is to increase accuracy and calibration compared to individual forecasts and to provide a single, cohesive prediction of future influenza. However, an ensemble may be composed of models that produce similar forecasts, causing issues with ensemble forecast performance and non-identifiability. To improve upon the above issues we propose a novel Cluster-Aggregate-Pool or `CAP' ensemble algorithm that first clusters together individual forecasts, aggregates individual models that belong to the same cluster into a single forecast (called a cluster forecast), and then pools together cluster forecasts via a linear pool. When compared to a non-CAP approach, we find that a CAP ensemble improves calibration by approximately 10% while maintaining similar accuracy to non-CAP alternatives. In addition, our CAP algorithm (i) generalizes past ensemble work associated with influenza forecasting and introduces a framework for future ensemble work, (ii) automatically accounts for missing forecasts from individual models, (iii) allows public health officials to participate in the ensemble by assigning individual models to clusters, and (iv) provide an additional signal about when peak influenza may be near."} {"id": "arxiv:2401.00085v1", "text": "This paper addresses the ``curse of dimensionality'' in the loss valuation of credit risk models. A dimension reduction methodology based on the Bayesian filter and smoother is proposed. This methodology is designed to achieve a fast and accurate loss valuation algorithm in credit risk modelling, but it can also be extended to valuation models of other risk types. The proposed methodology is generic, robust and can easily be implemented. Moreover, the accuracy of the proposed methodology in the estimation of expected loss and value-at-risk is illustrated by numerical experiments. The results suggest that, compared to the currently most used PCA approach, the proposed methodology provides more accurate estimation of expected loss and value-at-risk of a loss distribution."} {"id": "arxiv:2401.00090v1", "text": "In this paper, we address the problem of bounding conditional expectations when moment information of the underlying distribution and the random event conditioned upon are given. To this end, we propose an adapted version of the generalized moment problem which deals with this conditional information through a simple transformation. By exploiting conic duality, we obtain sharp bounds that can be used for distribution-free decision-making under uncertainty. Additionally, we derive computationally tractable mathematical programs for distributionally robust optimization (DRO) with side information by leveraging core ideas from ambiguity-averse uncertainty quantification and robust optimization, establishing a moment-based DRO framework for prescriptive stochastic programming."} {"id": "arxiv:2401.00094v2", "text": "The recent progress in language-based open-vocabulary object detection can be largely attributed to finding better ways of leveraging large-scale data with free-form text annotations. Training such models with a discriminative objective function has proven successful, but requires good positive and negative samples. However, the free-form nature and the open vocabulary of object descriptions make the space of negatives extremely large. Prior works randomly sample negatives or use rule-based techniques to build them. In contrast, we propose to leverage the vast knowledge built into modern generative models to automatically build negatives that are more relevant to the original data. Specifically, we use large-language-models to generate negative text descriptions, and text-to-image diffusion models to also generate corresponding negative images. Our experimental analysis confirms the relevance of the generated negative data, and its use in language-based detectors improves performance on two complex benchmarks. Code is available at \\url{https://github.com/xiaofeng94/Gen-Enhanced-Negs}."} {"id": "arxiv:2401.00099v1", "text": "Demagnetization in ferromagnetic transition metals driven by a femtosecond laser pulse is a fundamental problem in solid state physics, and its understanding is essential to the development of spintronics devices. Ab initio calculation of time-dependent magnetic moment in the velocity gauge so far has not been successful in reproducing the large amount of demagnetization observed in experiments. In this work, we propose a method to incorporate intraband transitions within the velocity gauge through a convective derivative in the crystal momentum space. Our results for transition-element bulk crystals (bcc Fe, hcp Co and fcc Ni) based on the time-dependent quantum Liouville equation show a dramatic enhancement in the amount of demagnetization after the inclusion of an intraband term, in agreement with experiments. We also find that the effect of intraband transitions to each ferromagnetic material is distinctly different because of their band structure and spin property differences. Our finding has a far-reaching impact on understanding of ultrafast demagnetization."} {"id": "arxiv:2401.00081v2", "text": "Synthetic data has made tremendous strides in various commercial settings including finance, healthcare, and virtual reality. We present a broad overview of prototypical applications of synthetic data in the financial sector and in particular provide richer details for a few select ones. These cover a wide variety of data modalities including tabular, time-series, event-series, and unstructured arising from both markets and retail financial applications. Since finance is a highly regulated industry, synthetic data is a potential approach for dealing with issues related to privacy, fairness, and explainability. Various metrics are utilized in evaluating the quality and effectiveness of our approaches in these applications. We conclude with open directions in synthetic data in the context of the financial domain."} {"id": "arxiv:2401.00078v3", "text": "Motivated by the question of how biological systems maintain homeostasis in changing environments, Shinar and Feinberg introduced in 2010 the concept of absolute concentration robustness (ACR). A biochemical system exhibits ACR in some species if the steady-state value of that species does not depend on initial conditions. Thus, a system with ACR can maintain a constant level of one species even as the environment changes. Despite a great deal of interest in ACR in recent years, the following basic question remains open: How can we determine quickly whether a given biochemical system has ACR? Although various approaches to this problem have been proposed, we show that they are incomplete. Accordingly, we present new methods for deciding ACR, which harness computational algebra. We illustrate our results on several biochemical signaling networks."} {"id": "arxiv:2401.00035v2", "text": "Complex dynamical systems are notoriously difficult to model because some degrees of freedom (e.g., small scales) may be computationally unresolvable or are incompletely understood, yet they are dynamically important. For example, the small scales of cloud dynamics and droplet formation are crucial for controlling climate, yet are unresolvable in global climate models. Semi-empirical closure models for the effects of unresolved degrees of freedom often exist and encode important domain-specific knowledge. Building on such closure models and correcting them through learning the structural errors can be an effective way of fusing data with domain knowledge. Here we describe a general approach, principles, and algorithms for learning about structural errors. Key to our approach is to include structural error models inside the models of complex systems, for example, in closure models for unresolved scales. The structural errors then map, usually nonlinearly, to observable data. As a result, however, mismatches between model output and data are only indirectly informative about structural errors, due to a lack of labeled pairs of inputs and outputs of structural error models. Additionally, derivatives of the model may not exist or be readily available. We discuss how structural error models can be learned from indirect data with derivative-free Kalman inversion algorithms and variants, how sparsity constraints enforce a \"do no harm\" principle, and various ways of modeling structural errors. We also discuss the merits of using non-local and/or stochastic error models. In addition, we demonstrate how data assimilation techniques can assist the learning about structural errors in non-ergodic systems. The concepts and algorithms are illustrated in two numerical examples based on the Lorenz-96 system and a human glucose-insulin model."} {"id": "arxiv:2401.00083v2", "text": "The Gouy phase is essential for accurately describing various wave phenomena, ranging from classical electromagnetic waves to matter waves and quantum optics. In this work, we employ phase-space methods based on the cross-Wigner transformation to analyze spatial and temporal interference in the evolution of matter waves characterized initially by a correlated Gaussian wave packet. First, we consider the cross-Wigner of the initial function with its free evolution, and second for the evolution through a double-slit arrangement. Different from the wave function which acquires a global Gouy phase, we find that the cross-Wigner acquires a Gouy phase difference due to different evolution times. The results suggest that temporal like-Gouy phases are important for an accurate description of temporal interference. Furthermore, we propose a technique based on the Wigner function to reconstruct the cross-Wigner from the spatial intensity interference term in a double-slit experiment with matter waves."} {"id": "arxiv:2401.00019v2", "text": "In this article, we discuss how a kind of hybrid computation, which employs symbolic, numeric, classic, and quantum algorithms, allows us to conduct Hartree-Fock electronic structure computation of molecules. In the proposed algorithm, we replace the Hartree-Fock equations with a set of equations composed of multivariate polynomials. We transform those polynomials to the corresponding Gröbner bases, and then we investigate the corresponding quotient ring, wherein the orbital energies, the LCAO coefficients, or the atomic coordinates are represented by the variables in the ring. In this quotient ring, the variables generate the transformation matrices that represent the multiplication with the monomial bases, and the eigenvalues of those matrices compose the roots of the equation. The quantum phase estimation (QPE) algorithm enables us to record those roots in the quantum states, which would be used in the input data for more advanced and more accurate quantum computations."} {"id": "arxiv:2401.00007v1", "text": "Epistemic emotions, such as curiosity and interest, drive the inquiry process. This study proposes a novel formulation of epistemic emotions such as curiosity and interest using two types of information gain generated by the principle of free energy minimization: Kullback-Leibler divergence(KLD) from Bayesian posterior to prior, which represents free energy reduction in recognition, and Bayesian surprise (BS), which represents the expected information gain by Bayesian prior update. By applying a Gaussian generative model with an additional uniform likelihood, we found that KLD and BS form an upward-convex function of surprise (minimized free energy and prediction error), similar to Berlyne's arousal potential functions, or the Wundt curve. We consider that the alternate maximization of BS and KLD generates an ideal inquiry cycle to approach the optimal arousal level with fluctuations in surprise, and that curiosity and interest drive to facilitate the cyclic process. We exhaustively analyzed the effects of prediction uncertainty (prior variance) and observation uncertainty (likelihood variance) on the peaks of the information gain function as optimal surprises. The results show that greater prediction uncertainty, meaning an open-minded attitude, and less observational uncertainty, meaning precise observation with attention, are expected to provide greater information gains through a greater range of exploration. The proposed mathematical framework unifies the free energy principle of the brain and the arousal potential theory to explain the Wundt curve as an information gain function and suggests an ideal inquiry process driven by epistemic emotions."} {"id": "arxiv:2401.00100v3", "text": "The classical Seiberg-Witten equations in dimensions three and four admit a natural generalization within a unified framework known as the generalized Seiberg-Witten (GSW) equations, which encompasses many important equations in gauge theory. This article proves that the averaged $L^2$-norm of any spinor with non-constant pointwise norm in the GSW equations on $\\mathbb R^4$ and $\\mathbb R^3$, measured over large-radius spheres, grows faster than a power of the radius, under a suitable curvature decay assumption. Separately, it is shown that if the Yang-Mills-Higgs energy of any solution of these equations is finite, then the pointwise norm of the spinor in it must converge to a non-negative constant at infinity. These two behaviors cannot occur simultaneously unless the spinor has constant pointwise norm. This work may be seen as partial generalization of results obtained by Taubes[Tau17a], and Nagy and Oliveira [NO19] for the Kapustin-Witten equations."} {"id": "arxiv:2401.00012v1", "text": "In this paper, we discuss the design and whole-cavity simulation of a Multi-Beam Inductive Output Tube (MBIOT) that uses a 3rd harmonic component to the drive voltage on the grid. High-efficiency inductive output tubes (IOTs) are typically characterized by efficiencies up to 70 - 75%. However, the achievement of efficiencies greater than 80% would substantially reduce the operating costs of next-generation accelerators. To achieve this goal, we consider the addition of a 3rd harmonic component to the drive signal on the grid. We anticipate that the MBIOT will be used to provide the rf power to drive RF linacs. We discuss and model an 8-beam MBIOT with a 700 MHz resonant cavity using beams with a voltage of 35 kV and an average current of 7.25 A yielding a perveance of about 1.1 microP. We simulate this MBIOT using the NEMESIS simulation code which has been extended using a three-dimensional Poisson solver based upon the Petsc package from Argonne National Laboratory. The effect of the 3rd harmonic on the efficiency is greatest when the phase of the 3rd harmonic is shifted by pi radians with respect to the fundamental drive signal and with 3rd harmonic powers greater than about 50% of the fundamental drive power. For the present example, we show that efficiencies approaching 82% are possible. Designs for the MBIOT input coupler, grids and output cavity have been developed based on these simulations and will be discussed."} {"id": "arxiv:2401.00021v1", "text": "Context: Students often misunderstand programming problem descriptions. This can lead them to solve the wrong problem, which creates frustration, obstructs learning, and imperils grades. Researchers have found that students can be made to better understand the problem by writing examples before they start programming. These examples are checked against correct and wrong implementations -- analogous to mutation testing -- provided by course staff. Doing so results in better student understanding of the problem as well as better test suites to accompany the program, both of which are desirable educational outcomes. Inquiry: Producing mutant implementations requires care. If there are too many, or they are too obscure, students will end up spending a lot of time on an unproductive task and also become frustrated. Instead, we want a small number of mutants that each correspond to common problem misconceptions. This paper presents a workflow with partial automation to produce mutants of this form which, notably, are not those produced by mutation-testing tools. Approach: We comb through student tests that fail a correct implementation. The student misconceptions are embedded in these failures. We then use methods to semantically cluster these failures. These clusters are then translated into conceptual mutants. These can then be run against student data to determine whether we they are better than prior methods. Some of these processes also enjoy automation. Knowledge: We find that student misconceptions illustrated by failing tests can be operationalized by the above process. The resulting mutants do much better at identifying student misconceptions. Grounding: Our findings are grounded in a manual analysis of student examples and a quantitative evaluation of both our clustering techniques and our process for making conceptual mutants. The clustering evaluation compares against a ground truth using standard cluster-correspondence measures, while the mutant evaluation examines how conceptual mutants perform against student data. Importance: Our work contributes a workflow, with some automation, to reduce the cost and increase the effectiveness of generating conceptually interesting mutants. Such mutants can both improve learning outcomes and reduce student frustration, leading to better educational outcomes. In the process, we also identify a variation of mutation testing not commonly discussed in the software literature."} {"id": "arxiv:2401.00022v1", "text": "Let $R$ be a finitely generated $\\mathbb N$-graded algebra domain over a Noetherian ring and let $I$ be a homogeneous ideal of $R$. Given $P\\in Ass(R/I)$ one defines the $v$-invariant $v_P(I)$ of $I$ at $P$ as the least $c\\in \\mathbb N$ such that $P=I:f$ for some $f\\in R_c$. A classical result of Brodmann asserts that $Ass(R/I^n)$ is constant for large $n$. So it makes sense to consider a prime ideal $P\\in Ass(R/I^n)$ for all the large $n$ and investigate how $v_P(I^n)$ depends on $n$. We prove that $v_P(I^n)$ is eventually a linear function of $n$. When $R$ is the polynomial ring over a field this statement has been proved independently also by Ficarra and Sgroi in a recent preprint."} {"id": "arxiv:2401.00028v3", "text": "The laws of model size, data volume, computation and model performance have been extensively studied in the field of Natural Language Processing (NLP). However, the scaling laws in Optical Character Recognition (OCR) have not yet been investigated. To address this, we conducted comprehensive studies that involved examining the correlation between performance and the scale of models, data volume and computation in the field of text recognition.Conclusively, the study demonstrates smooth power laws between performance and model size, as well as training data volume, when other influencing factors are held constant. Additionally, we have constructed a large-scale dataset called REBU-Syn, which comprises 6 million real samples and 18 million synthetic samples. Based on our scaling law and new dataset, we have successfully trained a scene text recognition model, achieving a new state-ofthe-art on 6 common test benchmarks with a top-1 average accuracy of 97.42%. The models and dataset are publicly available at https://github.com/large-ocr-model/large-ocr-model.github.io."} {"id": "arxiv:2401.00030v1", "text": "Despite existing for millennia, tuberculosis (TB) remains a persistent global health challenge. A significant obstacle in controlling TB spread is the need for a rapid, portable, sensitive, and accurate diagnostic test. Currently, sputum culture stands as a benchmark test for TB diagnosis. Although highly reliable, it necessitates advanced laboratory facilities and involves considerable testing time. In this context, we present a rapid, portable, and cost-effective optical fiber sensor designed to measure lipoarabinomannan (LAM), a TB biomarker found in patients' urine samples. Our sensing approach is based on the applications of phase shift-cavity ringdown spectroscopy (PS-CRDS) to an optical fiber cavity created by two fiber Bragg gratings. A tapered fiber is spliced inside the optical cavity to serve as the sensing head. We functionalize the tapered fiber surface with anti-LAM antigen CS-35 through a unique chemistry, creating a strong affinity for LAM molecules. We measure the phase difference between the cavity transmission and the reference modulating signal at the cavity output. The measured phase is directly proportional to the injected LAM concentrations in aqueous solutions over the sensing head. Our demonstrated sensor provides a detection limit of 10 pg/mL and a sensitivity of 2.6$^\\circ$/ng/mL. This sensor holds promise for numerous applications in the healthcare sector, particularly in low-resource settings."} {"id": "arxiv:2401.00043v1", "text": "Dark matter constitutes $26\\%$ of the total energy in our universe, but its nature remains elusive. Among the assortment of viable dark matter candidates, particles and fields with masses lighter than $40 \\mathrm{eV}$, called ultralight dark matter, stand out as particularly promising thanks to their feasible production mechanisms, consistency with current observations, and diverse and testable predictions. In light of ongoing and forthcoming experimental and observational efforts, it is important to advance the understanding of ultralight dark matter from theoretical and phenomenological perspectives: How does it interact with itself, ordinary matter, and gravity? What are some promising ways to detect it? In this thesis, we aim to explore the dynamics and interaction of ultralight dark matter and other astrophysically accessible hypothetical fields in a relatively model-independent way. Without making specific assumptions about their ultraviolet physics, we first demonstrate a systematic approach for constructing a classical effective field theory for both scalar and vector dark fields and discuss conditions for its validity. Then, we explore the interaction of ultralight dark fields, both gravitational and otherwise, within various contexts such as nontopological solitons, neutron stars, and gravitational waves."} {"id": "arxiv:2401.00051v1", "text": "The physical sciences require models tailored to specific nuances of different dynamics. In this work, we study outcome predictions in nuclear fusion tokamaks, where a major challenge are \\textit{disruptions}, or the loss of plasma stability with damaging implications for the tokamak. Although disruptions are difficult to model using physical simulations, machine learning (ML) models have shown promise in predicting these phenomena. Here, we first study several variations on masked autoregressive transformers, achieving an average of 5\\% increase in Area Under the Receiving Operating Characteristic metric above existing methods. We then compare transformer models to limited context neural networks in order to shed light on the ``memory'' of plasma effected by tokamaks controls. With these model comparisons, we argue for the persistence of a memory throughout the plasma \\textit{in the context of tokamaks} that our model exploits."} {"id": "arxiv:2401.00052v1", "text": "With the rapid evolution of Natural Language Processing (NLP), Large Language Models (LLMs) like ChatGPT have emerged as powerful tools capable of transforming various sectors. Their vast knowledge base and dynamic interaction capabilities represent significant potential in improving education by operating as a personalized assistant. However, the possibility of generating incorrect, biased, or unhelpful answers are a key challenge to resolve when deploying LLMs in an education context. This work introduces an innovative architecture that combines the strengths of ChatGPT with a traditional information retrieval based chatbot framework to offer enhanced student support in higher education. Our empirical evaluations underscore the high promise of this approach."} {"id": "arxiv:2401.00053v1", "text": "With this paper, we begin a series of studies of extremal problems for estimating distributions of martingale transforms of bounded martingales. The Bellman functions corresponding to such problems are pointwise minimal diagonally concave functions on a horizontal strip, satisfying certain given boundary conditions. We describe the basic structures that arise when constructing such functions and present a solution in the case of asymmetric boundary conditions and a sufficiently small width of the strip."} {"id": "arxiv:2401.00058v1", "text": "The text in the profile of those offering their properties in England in English and in Germany in German, are compared to explore whether trust is built, and privacy concerns are reduced in the same way. Six methods of building trust are used by the landlords: (1) the level of formality, (2) distance and proximity, (3) emotiveness and humor, (4) being assertive and passive aggressive, (5) conformity to the platform language style and terminology and (6) setting boundaries. Privacy concerns are not usually reduced directly as this is left to the platform. The findings indicate that language has a limited influence and the platform norms and habits are the biggest influence."} {"id": "arxiv:2401.00071v1", "text": "We apply the shifted composition rule -- an information-theoretic principle introduced in our earlier work [AC23] -- to establish shift Harnack inequalities for the Langevin diffusion. We obtain sharp constants for these inequalities for the first time, allowing us to investigate their relationship with other properties of the diffusion. Namely, we show that they are equivalent to a sharp \"local gradient-entropy\" bound, and that they imply curvature upper bounds in a compelling reflection of the Bakry-Emery theory of curvature lower bounds. Finally, we show that the local gradient-entropy inequality implies optimal concentration of the score, a.k.a. the logarithmic gradient of the density."} {"id": "arxiv:2401.00075v1", "text": "This article proposes a Reynolds number scaling of the required grid points to perform wall-modeled LES of turbulent flows encountering separation off a solid surface. Based on comparisons between the various time scales in a non-equilibrium (due to the action of an external pressure gradient) turbulent boundary layer, a simple definition of the near-wall ``under-equilibrium\" and ``out-of-equilibrium\" scales is put forward (where ``under-equilibrium\" refers to scales governed by a quasi-balance between the viscous and the pressure gradient terms). It is shown that the former length scale varies with Reynolds number as lp Re^(-2/3). The same scaling is obtained from a simplified Green's function solution of the Poisson equation in the vicinity of the separation point. A-priori analysis demonstrates that the resolution required to reasonably predict the wall-shear stress (for example, errors lower than approximately 10-15% in the entire domain) in several nonequilibrium flows is at least O(10) lp irrespective of the Reynolds number and the Clauser parameter. Further, a series of a-posteriori validation studies are performed to determine the accuracy of this scaling including the flow over the Boeing speed bump, Song-Eaton diffuser, Notre-Dame Ramp, and the backward-facing step. The results suggest that for these flows, scaling the computational grids () such that / lp is independent of the Reynolds number results in accurate predictions of flow separation at the same ``nominal\" grid resolution across different Reynolds numbers. Finally, it is suggested that in the vicinity of the separation and reattachment points, the grid-point requirements for wall-modeled large eddy simulations may scale as Re^4/3, which is more restrictive than the previously proposed flat-plate boundary layer-based estimates (Re1) of Choi and Moin (Phys. Fluids, 2012) and Yang and Griffin (Phys. Fluids, 2021)."} {"id": "arxiv:2401.00079v3", "text": "We prove that if $A$ is a computable Hopfian finitely presented structure, then $A$ has a computable $d$-$Σ_2$ Scott sentence if and only if the weak Whitehead problem for $A$ is decidable. We use this to infer that every hyperbolic group as well as any polycyclic-by-finite group has a computable $d$-$Σ_2$ Scott sentence, thus covering two main classes of finitely presented groups. Our proof also implies that every weakly Hopfian finitely presented group is strongly defined by its $\\exists^+$-types, a question which arose in a different context."} {"id": "arxiv:2401.00082v1", "text": "Large ensembles of stochastically evolving interacting particles describe phenomena in diverse fields including statistical physics, neuroscience, biology, and engineering. In such systems, the infinitesimal evolution of each particle depends only on its own state (or history) and the states (or histories) of neighboring particles with respect to an underlying, possibly random, interaction graph. While these high-dimensional processes are typically too complex to be amenable to exact analysis, their dynamics are quite well understood when the interaction graph is the complete graph. In this case, classical theorems show that in the limit as the number of particles goes to infinity, the dynamics of the empirical measure and the law of a typical particle coincide and can be characterized in terms of a much more tractable dynamical system of reduced dimension called the mean-field limit. In contrast, until recently not much was known about corresponding convergence results in the complementary case when the interaction graph is sparse (i.e., with uniformly bounded average degree). This article provides a brief survey of classical work and then describes recent progress on the sparse regime that relies on a combination of techniques from random graph theory, Markov random fields, and stochastic analysis. The article concludes by discussing ramifications for applications and posing several open problems."} {"id": "arxiv:2401.00086v1", "text": "This paper proposes a computational model for policy administration. As an organization evolves, new users and resources are gradually placed under the mediation of the access control model. Each time such new entities are added, the policy administrator must deliberate on how the access control policy shall be revised to reflect the new reality. A well-designed access control model must anticipate such changes so that the administration cost does not become prohibitive when the organization scales up. Unfortunately, past Access Control research does not offer a formal way to quantify the cost of policy administration. In this work, we propose to model ongoing policy administration in an active learning framework. Administration cost can be quantified in terms of query complexity. We demonstrate the utility of this approach by applying it to the evolution of protection domains. We also modelled different policy administration strategies in our framework. This allowed us to formally demonstrate that domain-based policies have a cost advantage over access control matrices because of the use of heuristic reasoning when the policy evolves. To the best of our knowledge, this is the first work to employ an active learning framework to study the cost of policy deliberation and demonstrate the cost advantage of heuristic policy administration."} {"id": "arxiv:2401.00009v3", "text": "In the wake of the latest trends of artificial intelligence (AI), there has been a resurgence of claims and questions about the Turing test and its value, which are reminiscent of decades of practical \"Turing\" tests. If AI were quantum physics, by now several \"Schrödinger's\" cats would have been killed. It is time for a historical reconstruction of Turing's beautiful thought experiment. This paper presents a wealth of evidence, including new archival sources, and gives original answers to several open questions about Turing's 1950 paper, including its relation with early AI."} {"id": "arxiv:2401.00039v3", "text": "Defects are common in physical systems with boundaries, impurities or extensive measurements. The interaction between bulk and defect can lead to rich physical phenomena. Defects in gapless phases of matter with conformal symmetry usually flow to a defect conformal field theory (dCFT). Understanding the universal properties of dCFTs is a challenging task. In this paper, we propose a computational strategy applicable to a line defect in arbitrary dimensions. Our main assumption is that the defect has a UV description in terms of a local modification of the Hamiltonian so that we can compute the overlap between low-energy eigenstates of a system with or without the defect insertion. We argue that these overlaps contain a wealth of conformal data, including the $g$-function, which is an RG monotonic quantity that distinguishes different dCFTs, the scaling dimensions of defect creation operators $Δ^{+0}_α$ and changing operators $Δ^{+-}_α$ that live on the intersection of different types of line defects, and various OPE coefficients. We apply this method to the fuzzy sphere regularization of 3D CFTs and study the magnetic line defect of the 3D Ising CFT. Using exact diagonalization and DMRG, we report the non-perturbative results $g=0.602(2),Δ^{+0}_0=0.108(5)$ and $Δ^{+-}_0=0.84(5)$ for the first time. We also obtain other OPE coefficients and scaling dimensions. Our results have significant physical implications. For example, they constrain the possible occurrence of spontaneous symmetry breaking at line defects of the 3D Ising CFT. Our method can be potentially applied to various other dCFTs, such as plane defects and Wilson lines in gauge theories."} {"id": "arxiv:2401.00047v3", "text": "The usual gravitational wave memory effect can be understood as a change in the separation of two initially comoving observers due to a burst of gravitational waves. Over the past few decades, a wide variety of other, \"persistent\" observables which measure permanent effects on idealized detectors have been introduced, each probing distinct physical effects. These observables can be defined in (regions of) any spacetime where there exists a notion of radiation, such as perturbation theory off of a fixed background, nonlinear plane wave spacetimes, or asymptotically flat spacetimes. Many of the persistent observables defined in the literature have only been considered in asymptotically flat spacetimes, and the perturbative nature of such calculations has occasionally obscured deeper relationships between these observables that hold more generally. The goal of this paper is to show how these more general results arise, and to do so we focus on two observables related to the separation between two, potentially accelerated observers. The first is the curve deviation, which is a natural generalization of the displacement memory, and also contains what this paper proposes to call drift memory (previously called \"subleading displacement memory\") and ballistic memory. The second is a relative proper time shift that arises between the two observers, either at second order in their initial separation and relative velocity, or in the presence of relative acceleration. The results of this paper are, where appropriate, entirely non-perturbative in the curvature of spacetime, and so could be used beyond leading order in asymptotically flat spacetimes."} {"id": "arxiv:2401.00096v3", "text": "Atomistic simulations of matter, especially those that leverage first-principles (ab initio) electronic structure theory, provide a microscopic view of the world, underpinning much of our understanding of chemistry and materials science. Over the last decade or so, machine-learned force fields have transformed atomistic modeling by enabling simulations of ab initio quality over unprecedented time and length scales. However, early ML force fields have largely been limited by: (i) the substantial computational and human effort of developing and validating potentials for each particular system of interest; and (ii) a general lack of transferability from one chemical system to the next. Here we show that it is possible to create a general-purpose atomistic ML model, trained on a public dataset of moderate size, that is capable of running stable molecular dynamics for a wide range of molecules and materials. We demonstrate the power of the MACE-MP-0 model - and its qualitative and at times quantitative accuracy - on a diverse set of problems in the physical sciences, including properties of solids, liquids, gases, chemical reactions, interfaces and even the dynamics of a small protein. The model can be applied out of the box as a starting or \"foundation\" model for any atomistic system of interest and, when desired, can be fine-tuned on just a handful of application-specific data points to reach ab initio accuracy. Establishing that a stable force-field model can cover almost all materials changes atomistic modeling in a fundamental way: experienced users get reliable results much faster, and beginners face a lower barrier to entry. Foundation models thus represent a step towards democratising the revolution in atomic-scale modeling that has been brought about by ML force fields."} {"id": "arxiv:2401.00063v1", "text": "Analyzing the geometry of correlation sets constrained by general causal structures is of paramount importance for foundational and quantum technology research. Addressing this task is generally challenging, prompting the development of diverse theoretical techniques for distinct scenarios. Recently, novel hybrid scenarios combining different causal assumptions within different parts of the causal structure have emerged. In this work, we extend a graph theoretical technique to explore classical, quantum, and no-signaling distributions in hybrid scenarios, where classical causal constraints and weaker no-signaling ones are used for different nodes of the causal structure. By mapping such causal relationships into an undirected graph we are able to characterize the associated sets of compatible distributions and analyze their relationships. In particular we show how with our method we can construct minimal Bell-like inequalities capable of simultaneously distinguishing classical, quantum, and no-signaling behaviors, and efficiently estimate the corresponding bounds. The demonstrated method will represent a powerful tool to study quantum networks and for applications in quantum information tasks."} {"id": "arxiv:2401.00068v2", "text": "The description of the $T\\bar{T}$ deformation in terms of two-dimensional gravity is analyzed from the Hamiltonian point of view, in a manner analogous to the ADM description of general relativity. We find that the Hamiltonian constraints of the theory imply relations between target-space momentum at finite volume which are equivalent to the $T\\bar{T}$ finite volume flow equations. This fully-quantum $T\\bar{T}$ result emerges already at the classical level within the gravitational theory. We exemplify the analysis for the case when the undeformed sector is a collection of $D-2$ free massless scalars, where it is shown that -- somewhat non-trivially -- the target-space two-dimensional Poincaré symmetry is extended to $D$ dimensions. The connection between canonical quantization of this constrained Hamiltonian system and previous path integral quantizations is also discussed. We extend our analysis to the ``gravitational'' description of $J\\bar{T}$-type deformations, where it is found that the flow equations obtained involve deformations that twist the spatial boundary conditions."} {"id": "arxiv:2401.00041v3", "text": "We present a new formulation for Yang-Mills scattering amplitudes in any number of dimensions and at any loop order, based on the same combinatorial and binary-geometric ideas in kinematic space recently used to give an all-order description of Tr $φ^3$ theory. We propose that in a precise sense the amplitudes for a suitably \"stringy\" form of these two theories are identical, up to a simple shift of kinematic variables. This connection is made possible by describing the amplitudes for $n$ gluons via a \"scalar scaffolding\", arising from the scattering of $2n$ colored scalars coming in $n$ distinct pairs of flavors fusing to produce the gluons. Fundamental properties of the \"$u$-variables\", describing the \"binary geometry\" for surfaces appearing in the topological expansion, magically guarantee that the kinematically shifted Tr $φ^3$ amplitudes satisfy the physical properties needed to be interpreted as scaffolded gluons. These include multilinearity, gauge invariance, and factorization on tree- and loop- level gluon cuts. Our \"stringy\" scaffolded gluon amplitudes coincide with amplitudes in the bosonic string for extra-dimensional gluon polarizations at tree-level, but differ (and are simpler) at loop-level. We provide many checks on our proposal, including matching non-trivial leading singularities through two loops. The simple counting problem underlying the $u$ variables autonomously \"knows\" about everything needed to convert colored scalar to gluon amplitudes, exposing a striking \"discovery\" of Yang-Mills amplitudes from elementary combinatorial ideas in kinematic space."} {"id": "arxiv:2401.00091v2", "text": "Disformal transformations of Friedmann-Lemaître-Robertson-Walker and Bianchi geometries are analyzed in the context of scalar-tensor gravity. Novel aspects discussed explicitly are the $3+1$ splitting, the effective fluid equivalent of the gravitational scalar, Bianchi models, stealth solutions, and de Sitter solutions with non-constant scalar field (which are signatures of scalar-tensor gravity). Both pure disformal transformations and more general ones are discussed, including those containing higher derivatives of the scalar field recently introduced in the literature."} {"id": "arxiv:2401.00025v3", "text": "Learning from demonstration is a powerful method for teaching robots new skills, and having more demonstration data often improves policy learning. However, the high cost of collecting demonstration data is a significant bottleneck. Videos, as a rich data source, contain knowledge of behaviors, physics, and semantics, but extracting control-specific information from them is challenging due to the lack of action labels. In this work, we introduce a novel framework, Any-point Trajectory Modeling (ATM), that utilizes video demonstrations by pre-training a trajectory model to predict future trajectories of arbitrary points within a video frame. Once trained, these trajectories provide detailed control guidance, enabling the learning of robust visuomotor policies with minimal action-labeled data. Across over 130 language-conditioned tasks we evaluated in both simulation and the real world, ATM outperforms strong video pre-training baselines by 80% on average. Furthermore, we show effective transfer learning of manipulation skills from human videos and videos from a different robot morphology. Visualizations and code are available at: \\url{https://xingyu-lin.github.io/atm}."} {"id": "arxiv:2401.00073v3", "text": "The strategy of pre-training a large model on a diverse dataset, then fine-tuning for a particular application has yielded impressive results in computer vision, natural language processing, and robotic control. This strategy has vast potential in adaptive control, where it is necessary to rapidly adapt to changing conditions with limited data. Toward concretely understanding the benefit of pre-training for adaptive control, we study the adaptive linear quadratic control problem in the setting where the learner has prior knowledge of a collection of basis matrices for the dynamics. This basis is misspecified in the sense that it cannot perfectly represent the dynamics of the underlying data generating process. We propose an algorithm that uses this prior knowledge, and prove upper bounds on the expected regret after $T$ interactions with the system. In the regime where $T$ is small, the upper bounds are dominated by a term that scales with either $\\texttt{poly}(\\log T)$ or $\\sqrt{T}$, depending on the prior knowledge available to the learner. When $T$ is large, the regret is dominated by a term that grows with $δT$, where $δ$ quantifies the level of misspecification. This linear term arises due to the inability to perfectly estimate the underlying dynamics using the misspecified basis, and is therefore unavoidable unless the basis matrices are also adapted online. However, it only dominates for large $T$, after the sublinear terms arising due to the error in estimating the weights for the basis matrices become negligible. We provide simulations that validate our analysis. Our simulations also show that offline data from a collection of related systems can be used as part of a pre-training stage to estimate a misspecified dynamics basis, which is in turn used by our adaptive controller."} {"id": "arxiv:2401.00010v1", "text": "Online recruitment platforms typically employ Person-Job Fit models in the core service that automatically match suitable job seekers with appropriate job positions. While existing works leverage historical or contextual information, they often disregard a crucial aspect: job seekers' social relationships in professional networks. This paper emphasizes the importance of incorporating professional networks into the Person-Job Fit model. Our innovative approach consists of two stages: (1) defining a Workplace Heterogeneous Information Network (WHIN) to capture heterogeneous knowledge, including professional connections and pre-training representations of various entities using a heterogeneous graph neural network; (2) designing a Contextual Social Attention Graph Neural Network (CSAGNN) that supplements users' missing information with professional connections' contextual information. We introduce a job-specific attention mechanism in CSAGNN to handle noisy professional networks, leveraging pre-trained entity representations from WHIN. We demonstrate the effectiveness of our approach through experimental evaluations conducted across three real-world recruitment datasets from LinkedIn, showing superior performance compared to baseline models."} {"id": "arxiv:2401.00031v2", "text": "Decision-making is a dynamic process requiring perception, memory, and reasoning to make choices and find optimal policies. Traditional approaches to decision-making suffer from sample efficiency and generalization, while large-scale self-supervised pretraining has enabled fast adaptation with fine-tuning or few-shot learning in language and vision. We thus argue to integrate knowledge acquired from generic large-scale self-supervised pretraining into downstream decision-making problems. We propose Pretrain-Then-Adapt pipeline and survey recent work on data collection, pretraining objectives and adaptation strategies for decision-making pretraining and downstream inference. Finally, we identify critical challenges and future directions for developing decision foundation model with the help of generic and flexible self-supervised pretraining."} {"id": "arxiv:2401.00056v1", "text": "The life cycle of most marine invertebrates includes a planktonic larval stage before metamorphosis to bottom-dwelling adulthood. During larval stage, ciliary-mediated activity enables feeding (capture unicellular algae) and transport of materials (oxygen) required for the larva's growth, development, and successful metamorphosis. Investigating the underlying hydrodynamics of these behaviors is valuable for addressing fundamental biological questions (e.g., phenotypic plasticity) and advancing engineering applications. In this work, we combined microfluidics and fluorescence microscopy as a miniaturized PIV (mPIV) to study ciliary-medicated hydrodynamics during suspension feeding in sand dollar larvae (Dendraster excentricus). First, we confirmed the approach's feasibility by examining the underlying hydrodynamics (vortex patterns) for low- and high-fed larvae. Next, ciliary hydrodynamics were tracked from 11 days post-fertilization (DPF) to 20 DPF for 21 low-fed larvae. Microfluidics enabled the examination of baseline activities (without external flow) and behaviors in the presence of environmental cues (external flow). A library of qualitative vortex patterns and quantitative hydrodynamics was generated and shared as a stand alone repository. Results from mPIV (velocities) were used to examine the role of ciliary activity in transporting materials (oxygen). Given the laminar flow and the viscosity-dominated environments surrounding the larvae, overcoming the diffusive boundary layer is critical for the organism's survival. Peclet number analysis for oxygen transport suggested that ciliary velocities help overcome the diffusion dominated transport (max Pe numbers between 30-60). Microfluidics serving as mPIV provided a scalable and accessible approach for investigating the ciliary hydrodynamics of marine organisms."} {"id": "arxiv:2401.00064v1", "text": "The centre of the Milky Way hosts a supermassive black hole of 4 million solar masses called Sagittarius A*. This object has been observed for more than 20 years in the near infrared. This has confirmed some effects of General Relativity. In addition, recurrent observations have made it possible to detect flares, i.e. a much larger flux than the average, with a variability of the order of 30 minutes to 1 hour. In 2018, GRAVITY, using the 4 large telescopes of the VLTI/ESO, observed an orbital motion of the source of these flares. This thesis focuses on the modelling of these flares, using models of varying complexity, including one based on the phenomenon of magnetic reconnection. The latter corresponds to an abrupt change in the magnetic configuration around the black hole, releasing a large amount of energy. We are also looking at the problem of polarisation in General Relativity."} {"id": "arxiv:2401.00074v1", "text": "Quantum plasmonics of few electrons that was reported a few years ago in photodoping experiments has received little or no attention in doped nanomaterials both theoretically and experimentally. There are no studies of quantum plasmonics of large quantum dots of the size of Bohr excitonic radius or less that take electronic and geometric structure of the quantum dots. In this work, we studied extensively the absorption spectra of a protype quantum dot of sizeable ZnO of approximately 3.0 nm in diameter doped with Ga, and Al in diluted limit with density functional theory (DFT) plus Hubbard corrections (U). The localized surface plasmon resonances (LSPRs) were then determined from the real part of the dielectric function by correlating the negative portion of it to resonant spectral lines. Our results show the sensitivity of the spectral lines to distribution of the dopants, the electronic structure of the dopants, the polarization of the electric field, and size of the quantum dots. Even though our findings are based on stoichiometrically simple ZnO, it shows that the DFT + U method with numerical atomic basis can be used at reasonable computational cost to study quantum plasmonics of few electrons taking into account electronic and geometric structures, which is missing currently. The method can be extended to study magneto-optics of diluted magnetic semiconductor QDs."} {"id": "arxiv:2401.00077v2", "text": "Scientists are increasingly leveraging advances in instruments, automation, and collaborative tools to scale up their experiments and research goals, leading to new bursts of discovery. Various scientific disciplines, including neuroscience, have adopted key technologies to enhance collaboration, reproducibility, and automation. Drawing inspiration from advancements in the software industry, we present a roadmap to enhance the reliability and scalability of scientific operations for diverse research teams tackling large and complex projects. We introduce a five-level Capability Maturity Model describing the principles of rigorous scientific operations in projects ranging from small-scale exploratory studies to large-scale, multi-disciplinary research endeavors. Achieving higher levels of operational maturity necessitates the adoption of new, technology-enabled methodologies, which we refer to as SciOps. This concept is derived from the DevOps methodologies that have revolutionized the software industry. SciOps involves digital research environments that seamlessly integrate computational, automation, and AI-driven efforts throughout the research cycle-from experimental design and data collection to analysis and dissemination, ultimately leading to closed-loop discovery. This maturity model offers a framework for assessing and improving operational practices in multidisciplinary research teams, guiding them towards greater efficiency and effectiveness in scientific inquiry."} {"id": "arxiv:2401.00080v1", "text": "Re-identifying participants in ultra-distance running competitions can be daunting due to the extensive distances and constantly changing terrain. To overcome these challenges, computer vision techniques have been developed to analyze runners' faces, numbers on their bibs, and clothing. However, our study presents a novel gait-based approach for runners' re-identification (re-ID) by leveraging various pre-trained human action recognition (HAR) models and loss functions. Our results show that this approach provides promising results for re-identifying runners in ultra-distance competitions. Furthermore, we investigate the significance of distinct human body movements when athletes are approaching their endurance limits and their potential impact on re-ID accuracy. Our study examines how the recognition of a runner's gait is affected by a competition's critical point (CP), defined as a moment of severe fatigue and the point where the finish line comes into view, just a few kilometers away from this location. We aim to determine how this CP can improve the accuracy of athlete re-ID. Our experimental results demonstrate that gait recognition can be significantly enhanced (up to a 9% increase in mAP) as athletes approach this point. This highlights the potential of utilizing gait recognition in real-world scenarios, such as ultra-distance competitions or long-duration surveillance tasks."} {"id": "arxiv:2401.00020v2", "text": "Natural Medicinal Materials (NMMs) have a long history of global clinical applications and a wealth of records and knowledge. Although NMMs are a major source for drug discovery and clinical application, the utilization and sharing of NMM knowledge face crucial challenges, including the standardized description of critical information, efficient curation and acquisition, and language barriers. To address these, we developed ShennongAlpha, an AI-driven sharing and collaboration platform for intelligent knowledge curation, acquisition, and translation. For standardized knowledge curation, the platform introduced a Systematic Nomenclature to enable accurate differentiation and identification of NMMs. More than fourteen thousand Chinese NMMs have been curated into the platform along with their knowledge. Furthermore, the platform pioneered chat-based knowledge acquisition, standardized machine translation, and collaborative knowledge updating. Together, our study represents the first major advance in leveraging AI to empower NMM knowledge sharing, which not only marks a novel application of AI for Science, but also will significantly benefit the global biomedical, pharmaceutical, physician, and patient communities."} {"id": "arxiv:2401.00014v1", "text": "The prediction of tumor progression and chemotherapy response has been recently tackled exploiting Tumor Infiltrating Lymphocytes (TILs) and the nuclear protein Ki67 as prognostic factors. Recently, deep neural networks (DNNs) have been shown to achieve top results in estimating Ki67 expression and simultaneous determination of intratumoral TILs score in breast cancer cells. However, in the last ten years the extraordinary progress induced by deep models proliferated at least as much as their resource demand. The exorbitant computational costs required to query (and in some cases also to store) a deep model represent a strong limitation in resource-limited contexts, like that of IoT-based applications to support healthcare personnel. To this end, we propose a resource consumption-aware DNN for the effective estimate of the percentage of Ki67-positive cells in breast cancer screenings. Our approach reduced up to 75% and 89% the usage of memory and disk space respectively, up to 1.5x the energy consumption, and preserved or improved the overall accuracy of a benchmark state-of-the-art solution. Encouraged by such positive results, we developed and structured the adopted framework so as to allow its general purpose usage, along with a public software repository to support its usage."} {"id": "arxiv:2401.00015v1", "text": "Growth in the penetration of renewable energy sources makes supply more uncertain and leads to an increase in the system imbalance. This trend, together with the single imbalance pricing, opens an opportunity for balance responsible parties (BRPs) to perform energy arbitrage in the imbalance settlement mechanism. To this end, we propose a battery control framework based on distributional reinforcement learning (DRL). Our proposed control framework takes a risk-sensitive perspective, allowing BRPs to adjust their risk preferences: we aim to optimize a weighted sum of the arbitrage profit and a risk measure while constraining the daily number of cycles for the battery. We assess the performance of our proposed control framework using the Belgian imbalance prices of 2022 and compare two state-of-the-art RL methods, deep Q learning and soft actor-critic. Results reveal that the distributional soft actor-critic method can outperform other methods. Moreover, we note that our fully risk-averse agent appropriately learns to hedge against the risk related to the unknown imbalance price by (dis)charging the battery only when the agent is more certain about the price."} {"id": "arxiv:2401.00018v1", "text": "The Bayesian reconstruction entropy is considered an alternative to the Shannon-Jaynes entropy, as it does not exhibit the asymptotic flatness characteristic of the Shannon-Jaynes entropy and obeys the scale invariance. It is commonly utilized in conjunction with the maximum entropy method to derive spectral functions from Euclidean time correlators produced by lattice QCD simulations. This study expands the application of the Bayesian reconstruction entropy to the reconstruction of spectral functions for Matsubara or imaginary-time Green's functions in quantum many-body physics. Furthermore, it extends the Bayesian reconstruction entropy to implement the positive-negative entropy algorithm, enabling the analytic continuations of matrix-valued Green's functions on an element-wise manner. Both the diagonal and off-diagonal components of the matrix-valued Green's functions are treated equally. Benchmark results for the analytic continuations of synthetic Green's functions indicate that the Bayesian reconstruction entropy, when combined with the preblur trick, demonstrates comparable performance to the Shannon-Jaynes entropy. Notably, it exhibits greater resilience to noises in the input data, particularly when the noise level is moderate."} {"id": "arxiv:2401.00023v2", "text": "Image-to-image translation has gained popularity in the medical field to transform images from one domain to another. Medical image synthesis via domain transformation is advantageous in its ability to augment an image dataset where images for a given class is limited. From the learning perspective, this process contributes to data-oriented robustness of the model by inherently broadening the model's exposure to more diverse visual data and enabling it to learn more generalized features. In the case of generating additional neuroimages, it is advantageous to obtain unidentifiable medical data and augment smaller annotated datasets. This study proposes the development of a CycleGAN model for translating neuroimages from one field strength to another (e.g., 3 Tesla to 1.5). This model was compared to a model based on DCGAN architecture. CycleGAN was able to generate the synthetic and reconstructed images with reasonable accuracy. The mapping function from the source (3 Tesla) to target domain (1.5 Tesla) performed optimally with an average PSNR value of 25.69 $\\pm$ 2.49 dB and an MAE value of 2106.27 $\\pm$ 1218.37."} {"id": "arxiv:2401.00032v1", "text": "Computational imaging~(CI) has been attracting a lot of interest in recent years for its superiority over traditional imaging in various applications. In CI systems, information is generally acquired in an encoded form and subsequently decoded via processing algorithms, which is quite in line with the information transmission mode of modern communication, and leads to emerging studies from the viewpoint of information optical imaging. Currently, one of the most important issues to be theoretically studied for CI is to quantitatively evaluate the fundamental ability of information acquisition, which is essential for both objective performance assessment and efficient design of imaging system. In this paper, by incorporating the Bayesian filtering paradigm, we propose a framework for CI that enables quantitative evaluation and design of the imaging system, and demonstate it based on ghost imaging. In specific, this framework can provide a quantitative evaluation on the acquired information through Fisher information and Cramér-Rao Lower Bound (CRLB), and the intrinsic performance of the imaging system can be accessed in real-time. With simulation and experiments, the framework is validated and compared with existing linear unbiased algorithms. In particular, the image retrieval can reach the CRLB. Furthermore, information-driven adaptive design for optimizing the information acquisition procedure is also achieved. By quantitative describing and efficient designing, the proposed framework is expected to promote the practical applications of CI techniques."} {"id": "arxiv:2401.00038v1", "text": "We investigate the embedding formalism in conjunction with the Mellin transform to determine tree-level gluon amplitudes in AdS/CFT. Detailed computations of three to five-point correlators are conducted, ultimately distilling what were previously complex results for five-point correlators into a more succinct and comprehensible form. We then proceed to derive a recursion relation applicable to a specific class of $n$-point gluon amplitudes. This relation is instrumental in systematically constructing amplitudes for a range of topologies. We illustrate its efficacy by specifically computing six to eight-point functions. Despite the complexity encountered in the intermediate steps of the recursion, the higher-point correlator is succinctly expressed as a polynomial in boundary coordinates, upon which a specific differential operator acts. Remarkably, we observe that these amplitudes strikingly mirror their counterparts in flat space, traditionally computed using standard Feynman rules. This intriguing similarity has led us to propose a novel dictionary: comprehensive rules that bridge AdS Mellin amplitudes with flat-space gluon amplitudes."} {"id": "arxiv:2401.00042v2", "text": "We show that the bosonic sector of the $N=(1,0),\\, 6D$ Salam-Sezgin gauged supergravity model possesses a $T$-duality symmetry upon a circle reduction to $D=5$. We then construct a simple magnetic rotating string solution with two equal angular momenta. Applying the $T$-duality transformation to this solution, we obtain the general boosted rotating dyonic black string solutions whose global structures and thermodynamic quantities are also analyzed. Owing to the fact that the solutions are not asymptotically flat, we find that there are two distinct globally-different non-extremal solutions with two different sets of thermal dynamic variables, with both satisfying the thermodynamic first law and the corresponding Small relations. However, their BPS limit becomes the same and we show that it preserves one quarter of supersymmetry by directly solving the corresponding Killing spinor equations."} {"id": "arxiv:2401.00044v1", "text": "The subsolar mass primordial black hole (PBH) attracts attention as robust evidence of its primordial origin against the astrophysical black hole. Not only with themselves, PBHs can also form binaries with ordinary astrophysical objects, catching them by gravitational wave (GW) bremsstrahlung. We discuss the detectability of the inspiral GWs from binaries consisting of a PBH and a white dwarf (WD) by using space-borne gravitational wave interferometers like DECIGO. The conservative assessment shows the expected event number in three years by DECIGO is $\\mathcal{O}(10^{-6})$ for $M_\\mathrm{PBH} \\sim 0.1M_\\odot$. Possible enhancement mechanisms of WD-PBH binary formation may amplify this event rate. We discuss how large enhancement associated with WDs is required to detect WD-PBH merger events without violating the existing constraints on the PBH-PBH merger by the ground-based detector."} {"id": "arxiv:2401.00050v2", "text": "Higher-order topological insulators in two spatial dimensions display fractional corner charges. While fractional charges in one dimension are known to be captured by a many-body bulk invariant, computed by the Resta formula, a many-body bulk invariant for higher-order topology and the corresponding fractional corner charges remains elusive despite several attempts. Inspired by recent work by Tada and Oshikawa, we propose a well-defined many-body bulk invariant for $C_n$ symmetric higher-order topological insulators, which is valid for both non-interacting and interacting systems. Instead of relating them to the bulk quadrupole moment as was previously done, we show that in the presence of $C_n$ rotational symmetry, this bulk invariant can be directly identified with quantized fractional corner charges. In particular, we prove that the corner charge is quantized as $e/n$ with $C_n$ symmetry, leading to a $\\mathbb{Z}_n$ classification for higher-order topological insulators in two dimensions."} {"id": "arxiv:2401.00054v1", "text": "The late time acceleration of the Universe has challenged contemporary cosmology since its discovery. General Relativity explains this phenomenon by introducing the cosmological constant, named the standard cosmological model ($Λ$CDM). However, the cosmological constant solution has several drawbacks that have led cosmologists to explore and propose alternative models to explain the late time acceleration of the Universe. These alternatives span from models of a dynamical dark fluid, known as dark energy, to models of large-scale modifications of the gravitational interaction, known as modified gravity. The current dissertation intends to show several ways to investigate late-time cosmology or to look at probable places for future investigations in order to shed more light on the dark sector of the Universe..."} {"id": "arxiv:2401.00055v1", "text": "Research on algorithmic recourse typically considers how an individual can reasonably change an unfavorable automated decision when interacting with a fixed decision-making system. This paper focuses instead on the online setting, where system parameters are updated dynamically according to interactions with data subjects. Beyond the typical individual-level recourse, the online setting opens up new ways for groups to shape system decisions by leveraging the parameter update rule. We show empirically that recourse can be improved when users coordinate by jointly computing their feature perturbations, underscoring the importance of collective action in mitigating adverse automated decisions."} {"id": "arxiv:2401.00062v1", "text": "A critical function of an organization is to foster the level of integration (coordination and cooperation) necessary to achieve its objectives. The need to coordinate and motivation to cooperate emerges from the myriad dependencies between an organization's members and their work. Therefore, to reason about solutions to coordination and cooperation problems requires a robust representation that includes the underlying dependencies. We find that such a representation remains missing from formal organizational models, and we leverage semantics to bridge this gap. Drawing on well-established organizational research and our extensive fieldwork with one of North America's largest municipalities, (1) we introduce an ontology, formalized in first-order logic, that operationalizes concepts like outcome, reward, and epistemic dependence, and their links to potential integration risks; and (2) present real-world applications of this ontology to analyze and support integration in complex government infrastructure projects. Our ontology is implemented and validated in both Z3 and OWL. Key features of our model include inferable dependencies, explainable coordination and cooperation risks, and actionable insights on how dependency structures within an organization can be altered to mitigate the risks. Conceptualizing real-world challenges like incentive misalignment, free-riding, and subgoal optimization in terms of dependency structures, our semantics-based approach represents a novel method for modelling and enhancing coordination and cooperation. Integrated within a decision-support system, our model may serve as an impactful aid for organizational design and effectiveness. More broadly, our approach underscores the transformative potential of semantics in deriving tangible, real-world value from existing organization theory."} {"id": "arxiv:2401.00067v1", "text": "Statistical Shape Modeling (SSM) is a quantitative method for analyzing morphological variations in anatomical structures. These analyses often necessitate building models on targeted anatomical regions of interest to focus on specific morphological features. We propose an extension to \\particle-based shape modeling (PSM), a widely used SSM framework, to allow shape modeling to arbitrary regions of interest. Existing methods to define regions of interest are computationally expensive and have topological limitations. To address these shortcomings, we use mesh fields to define free-form constraints, which allow for delimiting arbitrary regions of interest on shape surfaces. Furthermore, we add a quadratic penalty method to the model optimization to enable computationally efficient enforcement of any combination of cutting-plane and free-form constraints. We demonstrate the effectiveness of this method on a challenging synthetic dataset and two medical datasets."} {"id": "arxiv:2401.00087v2", "text": "Rollback recovery strategies are well-known in concurrent and distributed systems. In this context, recovering from unexpected failures is even more relevant given the non-deterministic nature of execution, which means that it is practically impossible to foresee all possible process interactions. In this work, we consider a message-passing concurrent programming language where processes interact through message sending and receiving, but shared memory is not allowed. In this context, we design a checkpoint-based rollback recovery strategy that does not need a central coordination. For this purpose, we extend the language with three new operators: check, commit, and rollback. Furthermore, our approach is purely asynchronous, which is an essential ingredient to developing a source-to-source program instrumentation implementing a rollback recovery strategy."} {"id": "arxiv:2401.00088v1", "text": "Inspired by the synthesis of the high-pressure Fm-3m LaH10 superconducting superhydride, systematic density functional theory (DFT) calculations are performed to study ternaries that could be derived from it by replacing two of the hydrogen atoms with boron or carbon and varying the identity of the electropositive element. Though many of the resulting alkali-metal and alkaline-earth MC2H8 phases are predicted to be dynamically stable at mild pressures, their superconducting critical temperatures (Tcs) are low because their metallicity results from the filling of an electride-like band. Substitution with a trivalent element leads to phases with substantial metal d-character at the Fermi level whose Tcs are typically above 40 K. Among the MB2H8 phases examined, KB2H8, RbB2H8 and CsB2H8 are predicted to be dynamically stable at very mild pressures, and their stability is rationalized by a DFT-Chemical Pressure analysis that elucidates the role of the M atom size. Quantum anharmonic effects strongly affect the properties of KB2H8, the highest predicted Tc compound, near 10 GPa, but molecular dynamics simulations reveal it would decompose below its Tc at this pressure. Nonetheless, at ca. 50 GPa KB2H8 is predicted to be thermally stable with a superconducting figure of merit surpassing that of the recently synthesized LaBeH8."} {"id": "arxiv:2401.00036v3", "text": "We introduce a novel generative model, the Discrete Distribution Networks (DDN), that approximates data distribution using hierarchical discrete distributions. We posit that since the features within a network inherently capture distributional information, enabling the network to generate multiple samples simultaneously, rather than a single output, may offer an effective way to represent distributions. Therefore, DDN fits the target distribution, including continuous ones, by generating multiple discrete sample points. To capture finer details of the target data, DDN selects the output that is closest to the Ground Truth (GT) from the coarse results generated in the first layer. This selected output is then fed back into the network as a condition for the second layer, thereby generating new outputs more similar to the GT. As the number of DDN layers increases, the representational space of the outputs expands exponentially, and the generated samples become increasingly similar to the GT. This hierarchical output pattern of discrete distributions endows DDN with unique properties: more general zero-shot conditional generation and 1D latent representation. We demonstrate the efficacy of DDN and its intriguing properties through experiments on CIFAR-10 and FFHQ. The code is available at https://discrete-distribution-networks.github.io/"} {"id": "arxiv:2401.00061v1", "text": "We introduce a physics-informed neural network (PINN) method to study thermoacoustic interactions leading to combustion instability in combustors. Specifically, we employ a PINN to investigate thermoacoustic interactions in a bluff body anchored flame combustor, representative of ramjet and industrial combustors. Vortex shedding and acoustic oscillations appear in such combustors, and their interactions lead to the phenomenon of vortex-acoustic lock-in. Acoustic pressure fluctuations at three locations and the total flame heat release rate serve as the measured data. The coupled parameterized model is based on the acoustic equations and the van der Pol oscillator for vortex shedding. The PINN was applied in the combustor, where the measurements suitable for a future machine learning application were not anticipated at the time of the experiments, as is the case in the vast majority of available data in the literature. We demonstrate a good performance of PINN in generating the acoustic field (pressure and velocity fluctuations) in the entire spatiotemporal domain, along with estimating all the parameters of the model. Therefore, this PINN-based model can potentially serve as an effective tool in improving existing combustors or designing new thermoacoustically stable and structurally efficient combustors."} {"id": "arxiv:2401.00040v1", "text": "We consider phenomenological aspects of a natural class of Standard Model-like supersymmetric F-theory vacua realized through flux breaking of rigid $E_7$ gauge factors. Three generations of Standard Model matter are realized in many of these vacua. We further find that many other Standard Model-like features are naturally compatible with these constructions. For example, dimension-4 and 5 terms associated with proton decay are ubiquitously suppressed. Many of these features are due to the group theoretical structure of $E_7$ and associated F-theory geometry. In particular, a set of approximate global symmetries descends from the $E_7$ group, leading to exponential suppression of undesired couplings."} {"id": "arxiv:2401.00060v3", "text": "Observatories need to measure and evaluate the scientific output and overall impact of their facilities. An observatory bibliography consists of the papers published using that observatory's data, typically gathered by searching the major journals for relevant keywords. Recently, the volume of literature and methods by which the publications pool is evaluated has increased. Efficient and standardized procedures are necessary to assign meaningful metadata; enable user-friendly retrieval; and provide the opportunity to derive reports, statistics, and visualizations to impart a deeper understanding of the research output. In 2021, a group of observatory bibliographers from around the world convened online to continue the discussions presented in Lagerstrom (2015). We worked to extract general guidelines from our experiences, techniques, and lessons learnt. The paper explores the development, application, and current status of telescope bibliographies and future trends. This paper briefly describes the methodologies employed in constructing databases, along with the various bibliometric techniques used to analyze and interpret them. We explain reasons for non-standardization and why it is essential for each observatory to identify metadata and metrics that are meaningful for them; caution the (over-)use of comparisons among facilities that are, ultimately, not comparable through bibliometrics; and highlight the benefits of telescope bibliographies, both for researchers within the astronomical community and for stakeholders beyond the specific observatories. There is tremendous diversity in the ways bibliographers track publications and maintain databases, due to parameters such as resources, type of observatory, historical practices, and reporting requirements to funders and outside agencies. However, there are also common sets of Best Practices."} {"id": "arxiv:2401.00002v1", "text": "We show evidence of particle acceleration at GEV energies associated directly with protons from the prompt emission of a long-duration M6-class solar flare on July 17, 2023, rather than from protons acceleration by shocks from its associated Coronal Mass Ejection (CME), which erupted with a speed of 1342 km/s. Solar Energetic Particles (SEP) accelerated by the blast have reached Earth, up to an almost S3 (strong) category of a radiation storm on the NOAA scale. Also, we show a temporal correlation between the fast rising of GOES-16 proton and muon excess at ground level in the count rate of the New-Tupi muon detector at the central SAA region. A Monte Carlo spectral analysis based on muon excess at New-Tupi is consistent with the acceleration of electrons and protons (ions) up to relativistic energies (GeV energy range) in the impulsive phase of the flare. In addition, we present another two marginal particle excesses (with low confidence) at ground-level detectors in correlation with the solar flare prompt emission."} {"id": "arxiv:2401.00097v3", "text": "This paper presents a regularized recursive identification algorithm with simultaneous on-line estimation of both the model parameters and the algorithms hyperparameters. A new kernel is proposed to facilitate the algorithm development. The performance of this novel scheme is compared with that of the recursive least squares algorithm in simulation."} {"id": "arxiv:2401.00092v3", "text": "We investigate the role of the spectral dimension $d_s$ in determining the universality of phase transitions on a complex network. Due to its structural heterogeneity, a complex network generally acts as a disordered system. Specifically, we study the synchronization and entrainment transitions in the nonequilibrium dynamics of the Kuramoto model and the phase transition of the equilibrium dynamics of the classical $XY$ model, thereby covering a broad spectrum from nonlinear dynamics to statistical and condensed matter physics. Using linear theory, we obtain a general relationship between the dynamics occurring on the network and the underlying network properties. This yields the lower critical spectral dimension of the phase synchronization and entrainment transitions in the Kuramoto model as $d_s=4$ and $d_s=2$ respectively, whereas for the phase transition in the $XY$ model it is $d_s=2$. To test our theoretical hypotheses, we employ a network where any two nodes on the network are connected with a probability proportional to a power law of the distance between the nodes; this realizes any desired $d_s\\in [1, \\infty)$. Our detailed numerical study agrees well with the prediction of linear theory for the phase synchronization transition in the Kuramoto model. However, it shows a clear entrainment transition in the Kuramoto model and phase transition in the $XY$ model at $d_s \\gtrsim 3$, not $d_s=2$ as predicted by linear theory. Our study indicates that network disorder in the region $2 \\leq d_s \\lesssim 3$ introduces strong finite-size fluctuations, which makes it extremely difficult to probe the existence of the ordered phase as predicted, affecting the dynamics profoundly."} {"id": "arxiv:2401.00057v1", "text": "Recent work on object-centric world models aim to factorize representations in terms of objects in a completely unsupervised or self-supervised manner. Such world models are hypothesized to be a key component to address the generalization problem. While self-supervision has shown improved performance however, OOD generalization has not been systematically and explicitly tested. In this paper, we conduct an extensive study on the generalization properties of contrastive world model. We systematically test the model under a number of different OOD generalization scenarios such as extrapolation to new object attributes, introducing new conjunctions or new attributes. Our experiments show that the contrastive world model fails to generalize under the different OOD tests and the drop in performance depends on the extent to which the samples are OOD. When visualizing the transition updates and convolutional feature maps, we observe that any changes in object attributes (such as previously unseen colors, shapes, or conjunctions of color and shape) breaks down the factorization of object representations. Overall, our work highlights the importance of object-centric representations for generalization and current models are limited in their capacity to learn such representations required for human-level generalization."} {"id": "arxiv:2401.00069v2", "text": "In this work we are motivated by factorization of bosonic quantum dynamics and we study the corresponding Lie algebras, which can potentially be infinite dimensional. To characterize such factorization, we identify conditions for these Lie algebras to be finite dimensional. We consider cases where each free Hamiltonian term is itself an element of the generated Lie algebra. In our approach, we develop new tools to systematically divide skew-hermitian bosonic operators into appropriate subspaces, and construct specific sequences of skew-hermitian operators that are used to gauge the dimensionality of the Lie algebras themselves. The significance of our result relies on conditions that constrain only the independently controlled generators in a particular Hamiltonian, thereby providing an effective algorithm for verifying the finiteness of the generated Lie algebra. In addition, our results are tightly connected to mathematical work where the polynomials of creation and annihilation operators are known as the Weyl algebra. Our work paves the way for better understanding factorization of bosonic dynamics relevant to quantum control and quantum technology."} {"id": "arxiv:2401.00013v1", "text": "We analyze a general problem in a crowd-sourced setting where one user asks a question (also called item) and other users return answers (also called labels) for this question. Different from existing crowd sourcing work which focuses on finding the most appropriate label for the question (the \"truth\"), our problem is to determine a ranking of the users based on their ability to answer questions. We call this problem \"ability discovery\" to emphasize the connection to and duality with the more well-studied problem of \"truth discovery\". To model items and their labels in a principled way, we draw upon Item Response Theory (IRT) which is the widely accepted theory behind standardized tests such as SAT and GRE. We start from an idealized setting where the relative performance of users is consistent across items and better users choose better fitting labels for each item. We posit that a principled algorithmic solution to our more general problem should solve this ideal setting correctly and observe that the response matrices in this setting obey the Consecutive Ones Property (C1P). While C1P is well understood algorithmically with various discrete algorithms, we devise a novel variant of the HITS algorithm which we call \"HITSNDIFFS\" (or HND), and prove that it can recover the ideal C1P-permutation in case it exists. Unlike fast combinatorial algorithms for finding the consecutive ones permutation (if it exists), HND also returns an ordering when such a permutation does not exist. Thus it provides a principled heuristic for our problem that is guaranteed to return the correct answer in the ideal setting. Our experiments show that HND produces user rankings with robustly high accuracy compared to state-of-the-art truth discovery methods. We also show that our novel variant of HITS scales better in the number of users than ABH, the only prior spectral C1P reconstruction algorithm."} {"id": "arxiv:2401.00004v1", "text": "The paper considers a non-reductionist theory of consciousness, which is not reducible to theories of reality and to physiological or psychological theories. Following D.I.Dubrovsky's \"informational approach\" to the \"Mind-Brain Problem\", we consider the reality through the prism of information about observed phenomena, which, in turn, is perceived by subjective reality through sensations, perceptions, feelings, etc., which, in turn, are information about the corresponding brain processes. Within this framework the following principle of the Information Theory of Consciousness (ITS) development is put forward: the brain discovers all possible causal relations in the external world and makes all possible inferences by them. The paper shows that ITS built on this principle: (1) also base on the information laws of the structure of external world; (2) explains the structure and functioning of the brain functional systems and cellular ensembles; (3) ensures maximum accuracy of predictions and the anticipation of reality; (4) resolves emerging contradictions and (5) is an information theory of the brain's reflection of reality."} {"id": "arxiv:2401.00008v1", "text": "Local Binary Patterns (LBP) are extensively used to analyze local texture features of an image. Several new extensions to LBP-based texture descriptors have been proposed, focusing on improving noise robustness by using different coding or thresholding schemes. In this paper we propose three algorithms (LBP), Shift Local Binary Pattern (SLBP), and Multi Shift Local Binary Pattern (MSLBP),to extract features for palmprint images that help to obtain the best unique and characteristic values of an image for identification. The Principal Component Analysis (PCA) algorithm has been applied to reduce the size of the extracted feature matrix in random space and in the matching process; the Linear Discriminant Analysis (LDA) algorithm is used. Several experiments were conducted on the large multispectral database (blue, green, red, and infrared) of the University of Hong Kong. As result, distinguished and high results were obtained where it was proved that, the blue spectrum is superior to all spectra perfectly."} {"id": "arxiv:2401.00016v2", "text": "The potential for augmenting the segmentation of brain tumors through the use of few-shot learning is vast. Although several deep learning networks (DNNs) demonstrate promising results in terms of segmentation, they require a substantial quantity of training data in order to produce suitable outcomes. Furthermore, a major issue faced by most of these models is their ability to perform well when faced with unseen classes. To address these challenges, we propose a one-shot learning model for segmenting brain tumors in magnetic resonance images (MRI) of the brain, based on a single prototype similarity score. Leveraging the recently developed techniques of few-shot learning, which involve the utilization of support and query sets of images for training and testing purposes, we strive to obtain a definitive tumor region by focusing on slices that contain foreground classes. This approach differs from other recent DNNs that utilize the entire set of images. The training process for this model is carried out iteratively, with each iteration involving the selection of random slices that contain foreground classes from randomly sampled data as the query set, along with a different random slice from the same sample as the support set. In order to distinguish the query images from the class prototypes, we employ a metric learning-based approach that relies on non-parametric thresholds. We employ the multimodal Brain Tumor Image Segmentation (BraTS) 2021 dataset, which comprises 60 training images and 350 testing images. The effectiveness of the model is assessed using the mean dice score and mean Intersection over Union (IoU) score."} {"id": "arxiv:2401.00027v2", "text": "Coarse-to-fine schemes are widely used in traditional single-image motion deblur; however, in the context of deep learning, existing multi-scale algorithms not only require the use of complex modules for feature fusion of low-scale RGB images and deep semantics, but also manually generate low-resolution pairs of images that do not have sufficient confidence. In this work, we propose a multi-scale network based on single-input and multiple-outputs(SIMO) for motion deblurring. This simplifies the complexity of algorithms based on a coarse-to-fine scheme. To alleviate restoration defects impacting detail information brought about by using a multi-scale architecture, we combine the characteristics of real-world blurring trajectories with a learnable wavelet transform module to focus on the directional continuity and frequency features of the step-by-step transitions between blurred images to sharp images. In conclusion, we propose a multi-scale network with a learnable discrete wavelet transform (MLWNet), which exhibits state-of-the-art performance on multiple real-world deblurred datasets, in terms of both subjective and objective quality as well as computational efficiency."} {"id": "arxiv:2401.00033v1", "text": "Design patterns provide a systematic way to convey solutions to recurring modeling challenges. This paper introduces design patterns for hybrid modeling, an approach that combines modeling based on first principles with data-driven modeling techniques. While both approaches have complementary advantages there are often multiple ways to combine them into a hybrid model, and the appropriate solution will depend on the problem at hand. In this paper, we provide four base patterns that can serve as blueprints for combining data-driven components with domain knowledge into a hybrid approach. In addition, we also present two composition patterns that govern the combination of the base patterns into more complex hybrid models. Each design pattern is illustrated by typical use cases from application areas such as climate modeling, engineering, and physics."} {"id": "arxiv:2401.00034v1", "text": "Integrating nanoscale opto-electronic functions is vital for applications such as optical emitters, detectors, and quantum information. Lanthanide atoms show great potential in this endeavor due to their intrinsic transitions. Here, we investigate Er adatoms on Si(100)-2x1 at 9K using a scanning tunneling microscope (STM) coupled to a tunable laser. Er adatoms display two main adsorption configurations that are optically excited between 800 nm and 1200 nm while the STM reads the resulting photocurrents. Our spectroscopic method reveals that various photocurrent signals stem from the bare silicon surface or Er adatoms. Additional photocurrent peaks appear as the signature of the Er adatoms relaxation, triggering efficient dissociation of nearby trapped excitons. Calculations using the density functional theory with spin-orbit coupling correction highlight the origin of the observed photocurrent peaks as specific 4f->4f or 4f->5d transitions. This spectroscopic technique can pave the way to an optoelectronic analysis of atomic and molecular assemblies by offering unique insight into their intrinsic quantum properties."} {"id": "arxiv:2401.00037v2", "text": "The tasks of designing RNAs are discrete optimization problems, and several versions of these problems are NP-hard. As an alternative to commonly used local search methods, we formulate these problems as continuous optimization and develop a general framework for this optimization based on a generalization of classical partition function which we call \"expected partition function\". The basic idea is to start with a distribution over all possible candidate sequences, and extend the objective function from a sequence to a distribution. We then use gradient descent-based optimization methods to improve the extended objective function, and the distribution will gradually shrink towards a one-hot sequence (i.e., a single sequence). As a case study, we consider the important problem of mRNA design with wide applications in vaccines and therapeutics. While the recent work of LinearDesign can efficiently optimize mRNAs for minimum free energy (MFE), optimizing for ensemble free energy is much harder and likely intractable. Our approach can consistently improve over the LinearDesign solution in terms of ensemble free energy, with bigger improvements on longer sequences."} {"id": "arxiv:2401.00045v1", "text": "In this work, we compare the SMBH and host galaxy properties of X-ray obscured and unobscured AGN. For that purpose, we use $\\sim 35 000$ X-ray detected AGN in the 4XMM-DR11 catalogue for which there are available measurements for their X-ray spectral parameters, from the XMM2Athena Horizon 2020 European project. We calculate the host galaxy properties via SED fitting analysis. Our final sample consists of 1 443 AGN. In the first part of our analysis, we use different N$_H$ thresholds (10$^{23}$ cm$^{-2}$ or 10$^{22}$ cm$^{-2}$), taking also into account the uncertainties associated with the N$_H$ measurements, to classify these sources into obscured and unobscured. We find that obscured AGN tend to live in more massive systems that have lower SFR compared to their unobscured counterparts. However, only the difference in stellar mass, M$_*$, appears statistically significant ($>2σ$). The results do not depend on the N$_H$ threshold used to classify AGN. The differences in M$_*$ and SFR are not statistically significant for luminous AGN ($\\rm log (L_{X,2-10 KeV}/erg s^{-1})> 44$). Our findings also show that unobscured AGN have, on average, higher specific black hole accretion rates compared to their obscured counterparts. In the second part of our analysis, we cross-match the 1 443 X-ray AGN with the SDSS DR16 quasar catalogue to obtain information on the SMBH properties of our sources. This results in 271 type 1 AGN, at $\\rm z<1.9$. Our findings show that type 1 AGN with increased N$_H$ ($>10^{22}$ cm$^{-2}$) tend to have higher M$_{BH}$ compared to AGN with lower N$_H$ values, at similar M$_*$. The M$_{BH}$/M$_*$ ratio remains consistent for N$_H$ values below 10$^{22}$ cm$^{-2}$, but it exhibits signs of an increase at higher N$_H$ values. Finally, we detect a correlation between $Γ$ and Eddington ratio, but only for type 1 sources with N$_H<10^{22}$ cm$^{-2}$."} {"id": "arxiv:2401.00048v1", "text": "We consider the massive scalar field equation $\\Box_{g_{RN}} φ= m^2 φ$ on any subextremal Reissner--Nordström exterior metric $g_{RN}$. We prove that solutions with localized initial data decay pointwise-in-time at the polynomial rate $t^{-\\frac{5}{6}+δ}$ in any spatially compact region (including the event horizon), for some small $ δ\\leq \\frac{1}{23} $. Moreover, assuming the validity of the Exponent Pair Conjecture on exponential sums in Number Theory, our result implies that decay upper bounds hold at the rate $t^{-\\frac{5}{6}+ε}$, for any arbitrarily small $ε>0$. In our previous work, we proved that each fixed angular mode decays at the exact rate $t^{-\\frac{5}{6}}$, thus the upper bound $t^{-\\frac{5}{6}+ε}$ is sharp, up to a $t^ε$ loss. Without the restriction to a fixed angular mode, the solution turns out to have an unbounded Fourier transform due to discrete frequencies associated to quasimodes, and caused by the occurrence of stable timelike trapping. Our analysis nonetheless shows that inverse-polynomial asymptotics in $t$ still hold after summing over all angular modes."} {"id": "arxiv:2401.00059v1", "text": "We examine the transmission of quantum particles (phonons, electrons, and photons) across interfaces, identifying universal patterns in diverse physical scenarios. Starting with classical wave equations, we quantize them and derive kinetic equations. Those are matching conditions for the distribution functions of particles at the interface. We note the time irreversibility of the derived kinetic equations -- an essential feature for accurately describing irreversible processes like heat transport. We identify the juncture in our derivation where the time symmetry of wave equations is disrupted, it is the assumption of the non-coherence of incident waves. Consequently, we infer that non-coherent transmission through the interface exhibits time irreversibility. We propose an experiment to validate this hypothesis."} {"id": "arxiv:2401.00070v1", "text": "Beineke, Harary and Ringel discovered a formula for the minimum genus of a torus in which the $n$-dimensional hypercube graph can be embedded. We give a new proof of the formula by building this surface as a union of certain faces in the hypercube's 2-skeleton. For odd dimension $n$, the entire 2-skeleton decomposes into $(n-1)/2$ copies of the surface, and the intersection of any two copies is the hypercube graph."} {"id": "arxiv:2401.00093v1", "text": "The rapid growth of the ride-hailing industry has revolutionized urban transportation worldwide. Despite its benefits, equity concerns arise as underserved communities face limited accessibility to affordable ride-hailing services. A key issue in this context is the vehicle rebalancing problem, where idle vehicles are moved to areas with anticipated demand. Without equitable approaches in demand forecasting and rebalancing strategies, these practices can further deepen existing inequities. In the realm of ride-hailing, three main facets of fairness are recognized: algorithmic fairness, fairness to drivers, and fairness to riders. This paper focuses on enhancing both algorithmic and rider fairness through a novel vehicle rebalancing method. We introduce an approach that combines a Socio-Aware Spatial-Temporal Graph Convolutional Network (SA-STGCN) for refined demand prediction and a fairness-integrated Matching-Integrated Vehicle Rebalancing (MIVR) model for subsequent vehicle rebalancing. Our methodology is designed to reduce prediction discrepancies and ensure equitable service provision across diverse regions. The effectiveness of our system is evaluated using simulations based on real-world ride-hailing data. The results suggest that our proposed method enhances both accuracy and fairness in forecasting ride-hailing demand, ultimately resulting in more equitable vehicle rebalancing in subsequent operations. Specifically, the algorithm developed in this study effectively reduces the standard deviation and average customer wait times by 6.48% and 0.49%, respectively. This achievement signifies a beneficial outcome for ride-hailing platforms, striking a balance between operational efficiency and fairness."} {"id": "arxiv:2401.00095v1", "text": "This paper presents a novel Automatic Essay Scoring (AES) algorithm tailored for the Portuguese-language essays of Brazil's Exame Nacional do Ensino Médio (ENEM), addressing the challenges in traditional human grading systems. Our approach leverages advanced deep learning techniques to align closely with human grading criteria, targeting efficiency and scalability in evaluating large volumes of student essays. This research not only responds to the logistical and financial constraints of manual grading in Brazilian educational assessments but also promises to enhance fairness and consistency in scoring, marking a significant step forward in the application of AES in large-scale academic settings."} {"id": "arxiv:2401.00072v2", "text": "We present a general framework for modeling power magnetic materials characteristics using deep neural networks. Magnetic materials represented by multidimensional characteristics (that mimic measurements) are used to train the neural autoencoder model in an unsupervised manner. The encoder is trying to predict the material parameters of a theoretical model, which is then used in a decoder part. The decoder, using the predicted parameters, reconstructs the input characteristics. The neural model is trained to capture a synthetically generated set of characteristics that can cover a broad range of material behaviors, leading to a model that can generalize on the underlying physics rather than just optimize the model parameters for a single measurement. After setting up the model, we prove its usefulness in the complex problem of modeling magnetic materials in the frequency and current (out-of-linear range) domains simultaneously, for which we use measured characteristics obtained for frequency up to $10$ MHz and H-field up to saturation."} {"id": "arxiv:2401.00049v3", "text": "Low-mass giants with large amounts of lithium (Li) have challenged stellar evolution for decades. One of the possibilities usually discussed to explain them involves the interaction with a close binary companion. This predicts that when compared against their non-enriched counterparts, Li-rich giants should preferentially be found as part of binary systems. In order to test this scenario, we assemble a sample of 1418 giants with radial velocities (RVs) from RAVE, GALAH, and Gaia, as well as stellar parameters and Li abundances from GALAH. Evolutionary states can be determined for 1030 of these giants. We develop a method that quantifies the degree of RV variability, which we use as a proxy for close binary companions. The method is tested and calibrated against samples of known RV standard stars and known spectroscopic binaries. We also compare the results of our RV variability analysis with binarity indicators from Gaia. We find that the accuracy of the classification is controlled by the precision of the RVs, which for the set of RVs available for the giants is 80-85%. Consistent with seismic studies, the resulting sample of giants contains a fraction of Li-rich objects in the red clump (RC) that is twice as large as that for first-ascent giants (RGB). Among RC giants, the fractions of Li-rich objects with high RV variability and with no RV variability are the same as those for Li-normal objects, which argues against a binary interaction scenario for the genesis of the bulk of Li-rich giants at that evolutionary stage. On the other hand, Li-rich giants in the RGB appear to have a small but detectable preference for higher RV variability, and thus possibly a larger close binary fraction, than the Li-normal giants at that stage. Additional measurements of the RVs of these giants at higher RV precision would greatly help confirm and more robustly quantify these results."} {"id": "arxiv:2401.00024v2", "text": "Achieving effective synergy between radiotherapy and immunotherapy is critical for optimizing tumor control and treatment outcomes. To explore the underlying mechanisms of this synergy, we have investigated a novel treatment approach known as personalized ultra-fractionated stereotactic adaptive radiation therapy (PULSAR), which emphasizes the impact of radiation timing on treatment efficacy. However, the precise mechanism remains unclear. Building on insights from small animal PULSAR studies, we developed a mathematical framework consisting of multiple ordinary differential equations to elucidate the temporal dynamics of tumor control resulting from radiation and the adaptive immune response. The model accounts for the migration and infiltration of T-cells within the tumor microenvironment. This proposed model establishes a causal and quantitative link between radiation therapy and immunotherapy, providing a valuable in-silico analysis tool for designing future PULSAR trials."} {"id": "arxiv:2401.00005v1", "text": "The work demonstrates that brain might reflect the external world causal relationships in the form of a logically consistent and prognostic model of reality, which shows up as consciousness. The paper analyses and solves the problem of statistical ambiguity and provides a formal model of causal relationships as probabilistic maximally specific rules. We suppose that brain makes all possible inferences from causal relationships. We prove that the suggested formal model has a property of an unambiguous inference: from consistent premises we infer a consistent conclusion. It enables a set of all inferences to form a consistent model of the perceived world. Causal relationships may create fixed points of cyclic inter-predictable properties. We consider the \"natural\" classification introduced by John St. Mill and demonstrate that a variety of fixed points of the objects' attributes forms a \"natural\" classification of the external world. Then we consider notions of \"natural\" categories and causal models of categories, introduced by Eleanor Rosch and Bob Rehder and demonstrate that fixed points of causal relationships between objects attributes, which we perceive, formalize these notions. If the \"natural\" classification describes the objects of the external world, and \"natural\" concepts the perception of these objects, then the theory of integrated information, introduced by G. Tononi, describes the information processes of the brain for \"natural\" concepts formation that reflects the \"natural\" classification. We argue that integrated information provides high accuracy of the objects identification. A computer-based experiment is provided that illustrates fixed points formation for coded digits."} {"id": "arxiv:2401.00098v1", "text": "In this work, we generalize the spacetime induced by a rotating cosmic string, taking into account anisotropic effects due the breaking of the Lorentz violation. In particular, we explore the energy levels of a massive spinless particle that is covariantly coupled to a uniform magnetic field aligned with the string. Subsequently, we introduce a scalar potential featuring both a Coulomb-type and a linear confining term and comprehensively solve the Klein-Gordon equations for each configuration. Finally, by imposing rigid-wall boundary conditions, we determine the Landau levels when the linear defect itself possesses magnetization. Notably, our analysis reveals the occurrence of Landau quantization even in the absence of gauge fields, provided the string possesses spin. Finally, the thermodynamic properties are computed as well in these scenarios."} {"id": "arxiv:2401.00089v4", "text": "Given two real symmetric matrices, their eigenvalue configuration is the relative arrangement of their eigenvalues on the real line. In this paper, we consider the following problem: given two parametric real symmetric matrices and an eigenvalue configuration, find a simple condition on the parameters such that their eigenvalues have the given configuration. We give an algorithm which expresses the eigenvalue configuration problem as a real root counting problem of certain symmetric polynomials, whose roots can be counted using the Fundamental Theorem of Symmetric Polynomials and Descartes' rule of signs."} {"id": "arxiv:2401.00003v6", "text": "Metamaterials with functional responses can exhibit varying properties under different conditions (e.g., wave-based responses or deformation-induced property variation). This work addresses the rapid inverse design of such metamaterials to meet target qualitative functional behaviors, a challenge due to its intractability and non-unique solutions. Unlike data-intensive and non-interpretable deep-learning-based methods, we propose the Random-forest-based Interpretable Generative Inverse Design (RIGID), a single-shot inverse design method for fast generation of metamaterial designs with on-demand functional behaviors. RIGID leverages the interpretability of a random forest-based \"design$\\rightarrow$response\" forward model, eliminating the need for a more complex \"response$\\rightarrow$design\" inverse model. Based on the likelihood of target satisfaction derived from the trained random forest, one can sample a desired number of design solutions using Markov chain Monte Carlo methods. We validate RIGID on acoustic and optical metamaterial design problems, each with fewer than 250 training samples. Compared to the genetic algorithm-based design generation approach, RIGID generates satisfactory solutions that cover a broader range of the design space, allowing for better consideration of additional figures of merit beyond target satisfaction. This work offers a new perspective on solving on-demand inverse design problems, showcasing the potential for incorporating interpretable machine learning into generative design under small data constraints."} {"id": "arxiv:2401.00166v1", "text": "This paper investigates block-level interference exploitation (IE) precoding for multi-user multiple-input single-output (MU-MISO) downlink systems. To overcome the need for symbol-level IE precoding to frequently update the precoding matrix, we propose to jointly optimize all the precoders or transmit signals within a transmission block. The resultant precoders only need to be updated once per block, and while not necessarily constant over all the symbol slots, we refer to the technique as block-level slot-variant IE precoding. Through a careful examination of the optimal structure and the explicit duality inherent in block-level power minimization (PM) and signal-to-interference-plus-noise ratio (SINR) balancing (SB) problems, we discover that the joint optimization can be decomposed into subproblems with smaller variable sizes. As a step further, we propose block-level slot-invariant IE precoding by adding a structural constraint on the slot-variant IE precoding to maintain a constant precoder throughout the block. A novel linear precoder for IE is further presented, and we prove that the proposed slot-variant and slot-invariant IE precoding share an identical solution when the number of symbol slots does not exceed the number of users. Numerical simulations demonstrate that the proposed precoders achieve a significant complexity reduction compared against benchmark schemes, without sacrificing performance."} {"id": "arxiv:2401.00167v1", "text": "Incorporating symmetry as an inductive bias into multi-agent reinforcement learning (MARL) has led to improvements in generalization, data efficiency, and physical consistency. While prior research has succeeded in using perfect symmetry prior, the realm of partial symmetry in the multi-agent domain remains unexplored. To fill in this gap, we introduce the partially symmetric Markov game, a new subclass of the Markov game. We then theoretically show that the performance error introduced by utilizing symmetry in MARL is bounded, implying that the symmetry prior can still be useful in MARL even in partial symmetry situations. Motivated by this insight, we propose the Partial Symmetry Exploitation (PSE) framework that is able to adaptively incorporate symmetry prior in MARL under different symmetry-breaking conditions. Specifically, by adaptively adjusting the exploitation of symmetry, our framework is able to achieve superior sample efficiency and overall performance of MARL algorithms. Extensive experiments are conducted to demonstrate the superior performance of the proposed framework over baselines. Finally, we implement the proposed framework in real-world multi-robot testbed to show its superiority."} {"id": "arxiv:2401.00171v1", "text": "We study the implementation of a Chebyshev spectral method with forward Euler integrator to investigate a peridynamic nonlocal formulation of Richards' equation. We prove the convergence of the fully-discretization of the model showing the existence and uniqueness of a solution to the weak formulation of the method by using the compactness properties of the approximated solution and exploiting the stability of the numerical scheme. We further support our results through numerical simulations, using initial conditions with different order of smoothness, showing reliability and robustness of the theoretical findings presented in the paper."} {"id": "arxiv:2401.00177v1", "text": "In this article, I will explore the nature of interference in translation, especially in technical and scientific texts, using a descriptivist approach. I will have a brief overview of the historical excursion of interference in technical and scientific translation. My aim is to explain this phenomenon and its causes with all its paradoxes, instead of simply condemning it as an example of supposedly bad translation. Thus, I will focus on its status in the bibliography of translation, on the motives for and consequences of interference in specialized translation, as well as on the nature of the arguments given for and against this phenomenon. Therefore the relationship between different societies has always been possible with the act of translation. When civilizations are examined throughout history, it is seen that the dissemination of knowledge among different societies has been achieved by translation. These societies have often become aware of the advancements in technology and science by means of translation. Therefore; translation becomes very significant in technical contact between societies and humans. Since the translation of technical texts is the preliminary scope of this thesis, it will be beneficial to have a brief look at the history of technical translation in the world."} {"id": "arxiv:2401.00178v1", "text": "We propose a methodology for the rheological characterization of a semisolid metal slurry using experimental squeeze flow data. The slurry is modeled as a structural thixotropic viscoplastic material, obeying the regularized Herschel-Bulkley constitutive equation. All rheological parameters are assumed to vary with the structure parameter that is governed by a first-order kinetics accounting for the material structure breakdown and build-up. The squeeze flow is simulated using finite elements in a Lagrangian framework. The evolution of the sample height has been studied for wide ranges of the Bingham and Reynolds numbers, the power-law exponent as well as the kinetics parameters of the structure parameter. Systematic comparisons have been carried out with available experimental data on a semisolid aluminium alloy (A356), where the sample is compressed from its topside under a specified strain of 80% at a temperature of 582 oC while the bottom side remains fixed. Excellent agreement with the experimental data could be achieved provided that at the initial instances (up to 0.01s) of the experiment the applied load is much higher than the nominal experimental load and that the yield stress and the power-law exponent vary linearly with the structure parameter. The first assumption implies that a different model, such as an elastoviscoplastic one, needs to be employed during the initial stages of the experiment. As for the second one, the evolution of the sample height can be reproduced allowing the yield stress to vary from 0 (no structure) to a maximum nominal value (full structure) and the power-law exponent from 0.2 to 1.4, i.e., from the shear-thinning to the shear-thickening regime. These variations are consistent with the internal microstructure variation pattern known to be exhibited by semisolid slurries."} {"id": "arxiv:2401.00116v1", "text": "Ionic liquids (ILs) are appealing electrolytes for their favorable physicochemical properties. However, despite their longstanding use, understanding the capacitive behavior of ILs remains challenging. This is largely due to the formation of a non-conventional electric double layer (EDL) at the electrode-electrolyte interface. This study shows that the short-range Yukawa interactions, representing the large anisotropically charged ILs, demix IL to create a spontaneous surface charge separation, which is reinforced by the strongly coupled charge interaction. The properties of the condensed layer, the onset of charge separation, and the rise of overscreening and crowding critically depend on the asymmetry of Yukawa interactions."} {"id": "arxiv:2401.00126v1", "text": "In this paper, we present a comprehensive investigation of stress propagation in a two-dimensional elastic circular disk. To accurately describe the displacements and stress fields within the disk, we employ a scalar and vector potential approach, representing them as sums of Bessel functions. The determination of the coefficients for these expansions is accomplished in the Laplace space, where we compare the boundary conditions. By converting the inverse Laplace transforms into complex integrals using residue calculus, we successfully derive explicit expressions for the displacements and stress fields. Notably, these expressions encompass primary, secondary, and surface waves, providing a thorough characterization of the stress propagation phenomena within the disk. Our findings contribute to the understanding of mechanical behavior in disk-shaped components and can be valuable in the design and optimization of such structures across various engineering disciplines."} {"id": "arxiv:2401.00135v1", "text": "Although sparse-view computed tomography (CT) has significantly reduced radiation dose, it also introduces severe artifacts which degrade the image quality. In recent years, deep learning-based methods for inverse problems have made remarkable progress and have become increasingly popular in CT reconstruction. However, most of these methods suffer several limitations: dependence on high-quality training data, weak interpretability, etc. In this study, we propose a fully unsupervised framework called Deep Radon Prior (DRP), inspired by Deep Image Prior (DIP), to address the aforementioned limitations. DRP introduces a neural network as an implicit prior into the iterative method, thereby realizing cross-domain gradient feedback. During the reconstruction process, the neural network is progressively optimized in multiple stages to narrow the solution space in radon domain for the under-constrained imaging protocol, and the convergence of the proposed method has been discussed in this work. Compared with the popular pre-trained method, the proposed framework requires no dataset and exhibits superior interpretability and generalization ability. The experimental results demonstrate that the proposed method can generate detailed images while effectively suppressing image artifacts.Meanwhile, DRP achieves comparable or better performance than the supervised methods."} {"id": "arxiv:2401.00143v1", "text": "Numerous systems require the capability to switch their operational modes seamlessly without any disruptions. The \"Synced Parallel Control Paths\" method is an innovative control system architecture designed for seamless mode switching. It features multiple parallel control paths: the primary path for essential operational references, and the auxiliary paths that continuously align with the primary, ensuring synchronized operation. This reduces operational disruptions common in traditional systems during mode transitions. This approach enhances system stability, reliability, and adaptability, making it a significant advancement in control technology for a variety of dynamic systems."} {"id": "arxiv:2401.00153v2", "text": "Inadequate generality across different organs and tasks constrains the application of ultrasound (US) image analysis methods in smart healthcare. Building a universal US foundation model holds the potential to address these issues. Nevertheless, the development of such foundational models encounters intrinsic challenges in US analysis, i.e., insufficient databases, low quality, and ineffective features. In this paper, we present a universal US foundation model, named USFM, generalized to diverse tasks and organs towards label efficient US image analysis. First, a large-scale Multi-organ, Multi-center, and Multi-device US database was built, comprehensively containing over two million US images. Organ-balanced sampling was employed for unbiased learning. Then, USFM is self-supervised pre-trained on the sufficient US database. To extract the effective features from low-quality US images, we proposed a spatial-frequency dual masked image modeling method. A productive spatial noise addition-recovery approach was designed to learn meaningful US information robustly, while a novel frequency band-stop masking learning approach was also employed to extract complex, implicit grayscale distribution and textural variations. Extensive experiments were conducted on the various tasks of segmentation, classification, and image enhancement from diverse organs and diseases. Comparisons with representative US image analysis models illustrate the universality and effectiveness of USFM. The label efficiency experiments suggest the USFM obtains robust performance with only 20% annotation, laying the groundwork for the rapid development of US models in clinical practices."} {"id": "arxiv:2401.00158v1", "text": "Question Answering over Knowledge Graph (KGQA) aims to seek answer entities for the natural language question from a large-scale Knowledge Graph~(KG). To better perform reasoning on KG, recent work typically adopts a pre-trained language model~(PLM) to model the question, and a graph neural network~(GNN) based module to perform multi-hop reasoning on the KG. Despite the effectiveness, due to the divergence in model architecture, the PLM and GNN are not closely integrated, limiting the knowledge sharing and fine-grained feature interactions. To solve it, we aim to simplify the above two-module approach, and develop a more capable PLM that can directly support subgraph reasoning for KGQA, namely ReasoningLM. In our approach, we propose a subgraph-aware self-attention mechanism to imitate the GNN for performing structured reasoning, and also adopt an adaptation tuning strategy to adapt the model parameters with 20,000 subgraphs with synthesized questions. After adaptation, the PLM can be parameter-efficient fine-tuned on downstream tasks. Experiments show that ReasoningLM surpasses state-of-the-art models by a large margin, even with fewer updated parameters and less training data. Our codes and data are publicly available at~\\url{https://github.com/RUCAIBox/ReasoningLM}."} {"id": "arxiv:2401.00193v1", "text": "In order to fully harness the potential of machine learning, it is crucial to establish a system that renders the field more accessible and less daunting for individuals who may not possess a comprehensive understanding of its intricacies. The paper describes the design of a system that integrates AutoML, XAI, and synthetic data generation to provide a great UX design for users. The system allows users to navigate and harness the power of machine learning while abstracting its complexities and providing high usability. The paper proposes two novel classifiers, Logistic Regression Forest and Support Vector Tree, for enhanced model performance, achieving 96\\% accuracy on a diabetes dataset and 93\\% on a survey dataset. The paper also introduces a model-dependent local interpreter called MEDLEY and evaluates its interpretation against LIME, Greedy, and Parzen. Additionally, the paper introduces LLM-based synthetic data generation, library-based data generation, and enhancing the original dataset with GAN. The findings on synthetic data suggest that enhancing the original dataset with GAN is the most reliable way to generate synthetic data, as evidenced by KS tests, standard deviation, and feature importance. The authors also found that GAN works best for quantitative datasets."} {"id": "arxiv:2401.00137v2", "text": "The extensive adoption of Self-supervised learning(SSL) has led to an increased security threat from backdoor attacks. While existing research has mainly focused on backdoor attacks in image classification, there has been limited exploration of their implications for object detection. Object detection plays a critical role in security-sensitive applications, such as autonomous driving, where backdoor attacks seriously threaten human life and property. In this work, we propose the first backdoor attack designed for object detection tasks in SSL scenarios, called Object Transform Attack (SSL-OTA). SSL-OTA employs a trigger capable of altering predictions of the target object to the desired category, encompassing two attacks: Naive Attack(NA) and Dual-Source Blending Attack (DSBA). NA conducts data poisoning during downstream fine-tuning of the object detector, while DSBA additionally injects backdoors into the pre-trained encoder. We establish appropriate metrics and conduct extensive experiments on benchmark datasets, demonstrating the effectiveness of our proposed attack and its resistance to potential defenses. Notably, both NA and DSBA achieve high attack success rates (ASR) at extremely low poisoning rates (0.5%). The results underscore the importance of considering backdoor threats in SSL-based object detection and contribute a novel perspective to the field."} {"id": "arxiv:2401.00103v2", "text": "We extend the notion of forward performance criteria to settings with random endowment in incomplete markets. Building on these results, we introduce and develop the novel concept of \\textit{forward optimized certainty equivalent (forward OCE)}, which offers a genuinely dynamic valuation mechanism that accommodates progressively adaptive market model updates, stochastic risk preferences, and incoming claims with arbitrary maturities. In parallel, we develop a new methodology to analyze the emerging stochastic optimization problems by directly studying the candidate optimal control processes for both the primal and dual problems. Specifically, we derive two new systems of forward-backward stochastic differential equations (FBSDEs) and establish necessary and sufficient conditions for optimality, and various equivalences between the two problems. This new approach is general and complements the existing one for forward performance criteria with random endowment based on backward stochastic partial differential equations (backward SPDEs) for the related value functions. We, also, consider representative examples for both forward performance criteria with random endowment and for forward OCE. Furthermore, for the case of exponential criteria, we investigate the connection between forward OCE and forward entropic risk measures."} {"id": "arxiv:2401.00196v1", "text": "In many causal studies, outcomes are censored by death, in the sense that they are neither observed nor defined for units who die. In such studies, the focus is usually on the stratum of always survivors up to a single fixed time s. Building on a recent strand of the literature, we propose an extended framework for the analysis of longitudinal studies, where units can die at different time points, and the main endpoints are observed and well defined only up to the death time. We develop a Bayesian longitudinal principal stratification framework, where units are cross classified according to the longitudinal death status. Under this framework, the focus is on causal effects for the principal strata of units that would be alive up to a time point s irrespective of their treatment assignment, where these strata may vary as a function of s. We can get precious insights into the effects of treatment by inspecting the distribution of baseline characteristics within each longitudinal principal stratum, and by investigating the time trend of both principal stratum membership and survivor-average causal effects. We illustrate our approach for the analysis of a longitudinal observational study aimed to assess, under the assumption of strong ignorability of treatment assignment, the causal effects of a policy promoting start ups on firms survival and hiring policy, where firms hiring status is censored by death."} {"id": "arxiv:2401.00110v7", "text": "Diffusion models without guidance generate very unrealistic samples. Guidance is used ubiquitously, and previous research has attributed its effect to low-temperature sampling that improves quality by trading off diversity. However, this perspective is incomplete. Our research shows that the choice of the loss objective is the underlying reason raw diffusion models fail to generate desirable samples. In this paper, (1) our analysis shows that the loss objective plays an important role in shaping the learned distribution and the MSE loss derived from theories holds assumptions that misalign with data in practice; (2) we explain the effectiveness of guidance methods from a new perspective of perceptual supervision; (3) we validate our hypothesis by training a diffusion model with a novel self-perceptual loss objective and obtaining much more realistic samples without the need for guidance. We hope our work paves the way for future explorations of the diffusion loss objective."} {"id": "arxiv:2401.00132v3", "text": "Multi-agent systems often require agents to collaborate with or compete against other agents with diverse goals, behaviors, or strategies. Agent modeling is essential when designing adaptive policies for intelligent machine agents in multiagent systems, as this is the means by which the ego agent understands other agents' behavior and extracts their meaningful policy representations. These representations can be used to enhance the ego agent's adaptive policy which is trained by reinforcement learning. However, existing agent modeling approaches typically assume the availability of local observations from other agents (modeled agents) during training or a long observation trajectory for policy adaption. To remove these constrictive assumptions and improve agent modeling performance, we devised a Contrastive Learning-based Agent Modeling (CLAM) method that relies only on the local observations from the ego agent during training and execution. With these observations, CLAM is capable of generating consistent high-quality policy representations in real-time right from the beginning of each episode. We evaluated the efficacy of our approach in both cooperative and competitive multi-agent environments. Our experiments demonstrate that our approach achieves state-of-the-art on both cooperative and competitive tasks, highlighting the potential of contrastive learning-based agent modeling for enhancing reinforcement learning."} {"id": "arxiv:2401.00187v4", "text": "When the structure deformation is dominated by the low-energy deformation mode, the structure hardens with the increase in the size (number of units) at small sizes. This anomalous behavior will eventually disappear with the decay length of the finite structure converging to a size-independent characteristic quantity, but the specific critical point at which the anomalous behavior disappears still cannot be accurately and concisely described. Here, under two steady states of the bistable chain, we observed anomalous size effects with constant and oscillating criticality (the proportion of inhomogeneous deformation), two criticalities exactly separate the increasing and decreasing intervals of stiffness variation. They are interrelated due to the implied symmetries between the two steady states. On the other hand, they are distinguished because of the opposite superposition modes under the two steady states. Specifically, the constant criticality corresponds to the anomalous size effect achieved by the competition mechanism, while the oscillating criticality reveals an anomalous size effect achieved by the new mechanism (cancellation mechanism). In the anomalous size effect achieved by the cancellation mechanism, the singular characteristics generated by the completely cancelled deformation make it very robust. This robustness reflects in that the anomalous effect is no longer limited to linear small deformation, but it can still be observed stably in nonlinear large deformation. Our study reinterprets the anomalous size effect at a quantitative level, and the proposed cancellation mechanism expands the possible application range of this anomalous effect."} {"id": "arxiv:2401.00107v1", "text": "Simulations of quantum systems in finite volume have proven to be a useful tool for calculating physical observables. Such studies to date have focused primarily on understanding the volume dependence of binding energies, from which it is possible to extract asymptotic properties of the corresponding bound state, as well as on extracting scattering information. For bound states, all properties depend on the size of the finite volume, and for precision studies it is important to understand such effects. In this work, we therefore derive the volume dependence of the mean squared radius of a two-body bound state, using a technique that can be generalized to other static properties in the future. We test our results with explicit numerical examples and demonstrate that we can robustly extract infinite-volume radii from finite-volume simulations in cubic boxes with periodic boundary conditions."} {"id": "arxiv:2401.00122v1", "text": "We develop a new efficient sequential approximate leverage score algorithm, SALSA, using methods from randomized numerical linear algebra (RandNLA) for large matrices. We demonstrate that, with high probability, the accuracy of SALSA's approximations is within $(1 + O({\\varepsilon}))$ of the true leverage scores. In addition, we show that the theoretical computational complexity and numerical accuracy of SALSA surpass existing approximations. These theoretical results are subsequently utilized to develop an efficient algorithm, named LSARMA, for fitting an appropriate ARMA model to large-scale time series data. Our proposed algorithm is, with high probability, guaranteed to find the maximum likelihood estimates of the parameters for the true underlying ARMA model. Furthermore, it has a worst-case running time that significantly improves those of the state-of-the-art alternatives in big data regimes. Empirical results on large-scale data strongly support these theoretical results and underscore the efficacy of our new approach."} {"id": "arxiv:2401.00140v2", "text": "We study supercritical age-structured branching models starting from a single particle with a random lifetime, where the reproduction law depends on the remaining lifetime of the parent. The lifespan of an individual is decided at its birth and its remaining lifetime decreases at the unit speed. A necessary and sufficient condition is provided for the convergence of the Malthusian normalized random measures. The Malthusian type limit theory in a functional form can be strengthened to hold with probability one under some ``$L\\log L$'' conditions. We further prove a central limit theory with a random normalization factor."} {"id": "arxiv:2401.00147v1", "text": "Suppose $G$ is a finitely generated infinite group, and $\\mathcal G$ is a graph of groups decomposition of $G$ such that the edge groups are finite. This paper establishes that the topology of the Floyd boundary of $G$ is uniquely determined by the topology of the Floyd boundary of each vertex group of $\\mathcal G$."} {"id": "arxiv:2401.00149v1", "text": "We construct a class of nonlinear coherent states (NLCSs) by introducing a more general nonlinear function and study their non-classical properties, specifically the second-order correlation function $g^{(2)}(0)$, Mandel parameter $Q$, squeezing, amplitude squared squeezing and Wigner function of the optical field. The results indicate that the non-classical properties of the new types of even and odd NLCSs crucially depend on nonlinear functions. More concretely, we find that the new even NLCSs could exhibit the photon-bunching effect whereas the new odd NLCSs could show photon-antibunching effect. The degree of squeezing is also significantly affected by the parameter selection of these NLCSs. By employing various forms of nonlinear functions, it becomes possible to construct NLCSs with diverse properties, thereby providing a theoretical foundation for corresponding experimental investigations."} {"id": "arxiv:2401.00150v1", "text": "We report the occurrence of vibrational resonance (VR) and the underlying mechanism in a simple piecewise linear electronic circuit, namely the Murali-Lakshmanan-Chua (MLC) circuit, driven by an additional biharmonic signal with widely different frequency. When the amplitude of the high-frequency force is tuned, the resultant vibrational resonance is used to detect the low-frequency signal and also to enhance it into a high-frequency signal. Further, we also show that even when the low-frequency signal is changed from sine wave to square and sawtooth waves, vibrational resonance can be used to detect and enhance them into high-frequency signals. These behaviors, confirmed by experimental results, are illustrated with appropriate analytical and numerical solutions of the corresponding circuit equations describing the system. Finally, we also verify the signal detection in the above circuit even with the addition of noise."} {"id": "arxiv:2401.00160v1", "text": "As indoor applications grow in diversity, wireless sensing, vital in areas like localization and activity recognition, is attracting renewed interest. Indoor wireless sensing relies on signal processing, particularly channel state information (CSI) based signal parameter estimation. Nonetheless, regarding reflected signals induced by dynamic human targets, no satisfactory algorithm yet exists for estimating the acceleration of dynamic path length change (DPLC), which is crucial for various sensing tasks in this context. Hence, this paper proposes DP-AcE, a CSI-based DPLC acceleration estimation algorithm. We first model the relationship between the phase difference of adjacent CSI measurements and the DPLC's acceleration. Unlike existing works assuming constant velocity, DP-AcE considers both velocity and acceleration, yielding a more accurate and objective representation. Using this relationship, an algorithm combining scaling with Fourier transform is proposed to realize acceleration estimation. We evaluate DP-AcE via the acceleration estimation and acceleration-based fall detection with the collected CSI. Experimental results reveal that, using distance as the metric, DP-AcE achieves a median acceleration estimation percentage error of 4.38%. Furthermore, in multi-target scenarios, the fall detection achieves an average true positive rate of 89.56% and a false positive rate of 11.78%, demonstrating its importance in enhancing indoor wireless sensing capabilities."} {"id": "arxiv:2401.00175v1", "text": "Distributed Computing in Blockchain Technology (BCT) hinges on a trust assumption among independent nodes. Without a third-party interface or what is known as a Blockchain Oracle, it can not interact with the external world. This Oracle plays a crucial role by feeding extrinsic data into the Blockchain, ensuring that Smart Contracts operate accurately in real time. The Oracle problem arises from the inherent difficulty in verifying the truthfulness of the data sourced by these Oracles. The genuineness of a Blockchain Oracle is paramount, as it directly influences the Blockchain's reliability, credibility, and scalability. To tackle these challenges, a strategy rooted in Byzantine fault tolerance φ is introduced. Furthermore, an autonomous system for sustainability and audibility, built on heuristic detection, is put forth. The effectiveness and precision of the proposed strategy outperformed existing methods using two real-world datasets, aimed to meet the authenticity standards for Blockchain Oracles."} {"id": "arxiv:2401.00183v1", "text": "In this paper we calculate the Belyi functions of the weighted trees of $(2,3)$-type with primitive special edge rotation groups. There are 21 Galois orbits of these trees: 6 rational orbits, 12 orbits over quadratic fields, and three orbits over fields of degree 3, 4 and 6. The highest degree of the calculated Belyi function is 32. The calculations are performed using modular functions techniques."} {"id": "arxiv:2401.00186v1", "text": "Linear Regression and neural networks are widely used to model data. Neural networks distinguish themselves from linear regression with their use of activation functions that enable modeling nonlinear functions. The standard argument for these activation functions is that without them, neural networks only can model a line. However, a novel explanation we propose in this paper for the impracticality of neural networks without activation functions, or linear neural networks, is that they actually reduce both training and testing performance. Having more parameters makes LNNs harder to optimize, and thus they require more training iterations than linear regression to even potentially converge to the optimal solution. We prove this hypothesis through an analysis of the optimization of an LNN and rigorous testing comparing the performance between both LNNs and linear regression on synthethic, noisy datasets."} {"id": "arxiv:2401.00191v1", "text": "In this paper, we introduce set-valued tensor complementarity problem where the elements of the involved tensors are defined based on a set-valued mapping. We study several properties of the solution set under the framework of set-valued mapping. We provide the necessary and sufficient conditions for the zero solution of a set-valued tensor complementarity problem. We introduce limit $R_0$-property for the set of tensors and establish a connection between limit $R_0$-property and the level boundedness of the merit function of the corresponding set-valued tensor complementarity problem."} {"id": "arxiv:2401.00197v1", "text": "Research into the prediction and analysis of perceived audio quality is hampered by the scarcity of openly available datasets of audio signals accompanied by corresponding subjective quality scores. To address this problem, we present the Open Dataset of Audio Quality (ODAQ), a new dataset containing the results of a MUSHRA listening test conducted with expert listeners from 2 international laboratories. ODAQ contains 240 audio samples and corresponding quality scores. Each audio sample is rated by 26 listeners. The audio samples are stereo audio signals sampled at 44.1 or 48 kHz and are processed by a total of 6 method classes, each operating at different quality levels. The processing method classes are designed to generate quality degradations possibly encountered during audio coding and source separation, and the quality levels for each method class span the entire quality range. The diversity of the processing methods, the large span of quality levels, the high sampling frequency, and the pool of international listeners make ODAQ particularly suited for further research into subjective and objective audio quality. The dataset is released with permissive licenses, and the software used to conduct the listening test is also made publicly available."} {"id": "arxiv:2401.00141v1", "text": "With rising concerns about the security of IoT devices, network operators need better ways to handle potential risks. Luckily, IoT devices show consistent patterns in how they communicate. But despite previous efforts, it remains unclear how knowledge of these patterns can be made available. As data marketplaces become popular in different domains, this paper1 proposes creating a special marketplace focused on IoT cybersecurity. The goal is to openly share knowledge about IoT devices' behavior, using structured data formats like Manufacturer Usage Description (MUD) files. To make this work, we employ technologies like blockchain and smart contracts to build a practical and secure foundation for sharing and accessing important information about how IoT devices should behave on the network. Our contributions are two-fold. (1) We identify the essential features of an effective marketplace for sharing data related to the expected behaviors of IoT devices. We develop a smart contract on the Ethereum blockchain with five concrete functions; and, (2) We implement a prototype of our marketplace in a private chain environment-our codes are publicly released. We demonstrate how effectively our marketplace functions through experiments involving MUD files from consumer IoT devices. Our marketplace enables suppliers and consumers to share MUD data on the Ethereum blockchain for under a hundred dollars, promoting accessibility and participation."} {"id": "arxiv:2401.00156v4", "text": "Radical subgroups play an important role in both finite group theory and representation theory. This is the first of a series of papers of ours in classifying radical $p$-subgroups of finite reductive groups and in verifying the inductive blockwise Alperin weight condition for them, contributing to the program of proving the Alperin weight conjecture by verifying its inductive condition for finite simple groups. In this paper we present a uniform method for classifying radical subgroups of finite reductive groups. As applications, we investigate the inductive blockwise Alperin weight condition for classical groups, as well as groups of type $F_4$."} {"id": "arxiv:2401.00112v2", "text": "This study presents an industry experience showcasing a vessel operational anomaly detection approach that utilizes semi-supervised deep learning models augmented with lightweight interpretable surrogate models, applied to an industrial sensorized vessel, called TUCANA. We leverage standard and Long Short-Term Memory (LSTM) autoencoders trained on normal operational data and tested with real anomaly-revealing data. We then provide a projection of the inference results on a lower-dimension data map generated by t-distributed stochastic neighbor embedding (t-SNE), which serves as an unsupervised baseline and shows the distribution of the identified anomalies. We also develop lightweight surrogate models using random forest and decision tree to promote transparency and interpretability for the inference results of the deep learning models and assist the engineer with an agile assessment of the flagged anomalies. The approach is empirically evaluated using real data from TUCANA. The empirical results show higher performance of the LSTM autoencoder -- as the anomaly detection module with effective capturing of temporal dependencies in the data -- and demonstrate the practicality of the lightweight surrogate models in providing helpful interpretability, which leads to higher efficiency for the engineer's decision-making."} {"id": "arxiv:2401.00104v2", "text": "Reinforcement learning (RL) is a powerful technique for training intelligent agents, but understanding why these agents make specific decisions can be quite challenging. This lack of transparency in RL models has been a long-standing problem, making it difficult for users to grasp the reasons behind an agent's behaviour. Various approaches have been explored to address this problem, with one promising avenue being reward decomposition (RD). RD is appealing as it sidesteps some of the concerns associated with other methods that attempt to rationalize an agent's behaviour in a post-hoc manner. RD works by exposing various facets of the rewards that contribute to the agent's objectives during training. However, RD alone has limitations as it primarily offers insights based on sub-rewards and does not delve into the intricate cause-and-effect relationships that occur within an RL agent's neural model. In this paper, we present an extension of RD that goes beyond sub-rewards to provide more informative explanations. Our approach is centred on a causal learning framework that leverages information-theoretic measures for explanation objectives that encourage three crucial properties of causal factors: causal sufficiency, sparseness, and orthogonality. These properties help us distill the cause-and-effect relationships between the agent's states and actions or rewards, allowing for a deeper understanding of its decision-making processes. Our framework is designed to generate local explanations and can be applied to a wide range of RL tasks with multiple reward channels. Through a series of experiments, we demonstrate that our approach offers more meaningful and insightful explanations for the agent's action selections."} {"id": "arxiv:2401.00172v1", "text": "We consider the estimation of small probabilities or other risk quantities associated with rare but catastrophic events. In the model-based literature, much of the focus has been devoted to efficient Monte Carlo computation or analytical approximation assuming the model is accurately specified. In this paper, we study a distinct direction on the propagation of model uncertainty and how it impacts the reliability of rare-event estimates. Specifically, we consider the basic setup of the exceedance of i.i.d. sum, and investigate how the lack of tail information of each input summand can affect the output probability. We argue that heavy-tailed problems are much more vulnerable to input uncertainty than light-tailed problems, reasoned through their large deviations behaviors and numerical evidence. We also investigate some approaches to quantify model errors in this problem using a combination of the bootstrap and extreme value theory, showing some positive outcomes but also uncovering some statistical challenges."} {"id": "arxiv:2401.00176v1", "text": "The paper is an attempt to apply the theory of dessins d'enfants to the theory of fullerenes. The classical results concerning the calculation of the dodecahedron Belyi function are presented and then applied to the calculation of the Belyi function of the barrel, and the euclidean geometry of the latter is investigated. The non-existence of the fullerene with the only hexagonal face is established by the methods of dessins d'enfants."} {"id": "arxiv:2401.00101v1", "text": "Knowledge of exact analytical functional forms for the pair correlation function $g_2(r)$ and its corresponding structure factor $S(k)$ of disordered many-particle systems is limited. For fundamental and practical reasons, it is highly desirable to add to the existing data base of analytical functional forms for such pair statistics. Here, we design a plethora of such pair functions in direct and Fourier spaces across the first three Euclidean space dimensions that are realizable by diverse many-particle systems with varying degrees of correlated disorder across length scales, spanning a wide spectrum of hyperuniform, typical nonhyperuniform and antihyperuniform ones. This is accomplished by utilizing an efficient inverse algorithm that determines equilibrium states with up to pair interactions at positive temperature that precisely match targeted forms for both $g_2(r)$ and $S(k)$. Among other results, we realize an example with the strongest hyperuniform property among known positive-temperature equilibrium states, critical-point systems (implying unusual 1D systems with phase transitions) that are not in the Ising universality class, systems that attain self-similar pair statistics under Fourier transformation, and an experimentally feasible polymer model. We show that our pair functions enable one to achieve systems with a wide range of translational order and self-diffusion coefficients $\\cal D$, which are inversely related to one another. One can design other realizable pair statistics via linear combinations of our functions or by applying our inverse procedure to other desirable functional forms. Our approach facilitates the inverse design of materials with desirable physical and chemical properties by tuning their pair statistics."} {"id": "arxiv:2401.00106v1", "text": "It is our purpose to study complete space-like self-expanders in the Minkovski space. By use of maximum principle of Omori-Yau type, we can obtain the rigidity theorems on $n$-dimensional complete space-like self-expanders in the Minkovski space $\\mathbb R^{n+1}_{1}$. For complete space-like self-expanders of dimension $2$, we give a classification of them under assumption of constant squared norm of the second fundamental form."} {"id": "arxiv:2401.00114v1", "text": "Both humans and social animals live in groups and are frequently faced to choose between options with different qualities. When no leader agents are controlling the group decision, consensus can be achieved through repeated interactions among group members. Various studies on CDM illustrate how the dynamics of opinions are determined by the structure of the social network and the methods that individuals use to share and update their opinion upon a social interaction. In this paper, we are interested in further exploring how cognitive, social, and environmental factors interactively contribute to determining the outcome of a collective best-of-n decision process involving asymmetric options, i.e., different costs and/or benefits for each option. We propose and study a novel model capturing those different factors, i) the error in processing social information, ii) the number of zealots (i.e., asocial agents who never change their opinion), iii) the option qualities, iv) the social connectivity structure, and v) the degree centrality of the asocial agents. By using the HMF approach, we study the impact of the above-mentioned factors in the decision dynamics. Our findings indicate that when susceptible agents use the voter model as a mechanism to update their opinion, both the number and the degree of connectivity of the zealots can lead the population to converge towards the lowest quality option. Instead, when susceptible agents use methods more cognitively demanding, the group is marginally impacted by the presence of zealots. The results of the analytical model are complemented and extended by agent-based simulations. Our analysis also shows that the network topology can modulate the influence of zealots on group dynamics."} {"id": "arxiv:2401.00117v1", "text": "We present an X-ray and UV investigation of five X-ray flares detected on two active systems, CC Eri and AB Dor, using the AstroSat observatory. The peak X-ray luminosities of the flares in the 0.3$-$7.0 keV band are found to be within 10$^{31-33}$ erg s$^{-1}$. Preliminary spectral analysis indicates the presence of three and four-temperature corona for CC Eri and AB Dor, respectively, where the highest temperature is found to vary with flare. The flare temperatures peaked at 51$-$59 MK for CC Eri and 29$-$44 MK for AB Dor. The peak emission measures of the flaring loops are estimated to be $\\sim$10$^{54}$ for CC Eri and $\\sim$10$^{55}$ cm$^{-3}$ for AB Dor. Global metallic abundances were also found to increase during flares."} {"id": "arxiv:2401.00118v1", "text": "Understanding the origin of electron incoherence is the first step toward a theoretical description of the non-Fermi liquid behavior of the high-T$_{c}$ cuprate superconductors. Such electron incoherence manifests itself most evidently in the non-Drude behavior of the optical response of the system and the anomalous density fluctuation behavior in the long wave length limit. The spectral weight transfer related to such dissipative response, which is absent in conventional Fermi liquid metal, has direct consequence on the dc transport property of the system in the normal state and the superfluid stiffness in the superconducting state. It is found that such electron incoherence remains significant even in the clean limit and at low temperature and thus must be attributed to the strong electron correlation effect in the cuprate superconductors. Here we study such an intrinsic effect in the 2D $t-J$ model through the variational calculation of its optical conductivity $σ(ω)$. We assume a resonating valence bond ground state as our starting point and find that a significant portion of the total optical spectral weight remains incoherent throughout the phase diagram. The optical absorption is found to extend all the way to an energy of the order of the bare band width. We find that both the total optical weight $\\bar{K}$ and the integrated incoherent optical weight $I$ increase monotonically with doping, with their ratio $R_{incoh}=I/\\bar{K}$ decreasing monotonically with doping. Our results indicate that the majority part of electron incoherence in the 2D $t-J$ model can be attributed to the electron fractionalization mechanism assumed in such a treatment. We also find that the Drude weight deduced from $D=\\bar{K}-I$ scales linearly with hole doping, without any sign of a non-monotonic behavior in the overdoped regime."} {"id": "arxiv:2401.00119v1", "text": "A Christ-Kiselev maximal theorem is proved for linear operators between quasi-Banach function lattices satisfying certain lattice geometrical conditions. The result is further explored for weighted Lorentz spaces, classical Lorentz spaces, and Wiener amalgams of Lebesgue function and sequence spaces. Extensions are made to Köthe dual operators and to operators on interpolation spaces of quasi-Banach function lattices. Several applications to maximal Fourier operators are presented."} {"id": "arxiv:2401.00120v1", "text": "We employ the eigen microstate approach to explore the self-organized criticality (SOC) in two celebrated sandpile models, namely, the BTW model and the Manna model. In both models, phase transitions from the absorbing-state to the critical state can be understood by the emergence of dominant eigen microstates with significantly increased weights. Spatial eigen microstates of avalanches can be uniformly characterized by a linear system size rescaling. The first temporal eigen microstates reveal scaling relations in both models. Furthermore, by finite-size scaling analysis of the first eigen microstate, we numerically estimate critical exponents i.e., $\\sqrt{σ_0 w_1}/\\tilde{v}_{1} \\propto L^D$ and $\\tilde{v}_{1} \\propto L^{D(1-τ_s)/2}$. Our findings could provide profound insights into eigen states of the universality and phase transition in non-equilibrium complex systems governed by self-organized criticality."} {"id": "arxiv:2401.00128v1", "text": "Glioblastoma (GBM) is one of the most aggressive and lethal human cancers. Intra-tumoral genetic heterogeneity poses a significant challenge for treatment. Biopsy is invasive, which motivates the development of non-invasive, MRI-based machine learning (ML) models to quantify intra-tumoral genetic heterogeneity for each patient. This capability holds great promise for enabling better therapeutic selection to improve patient outcomes. We proposed a novel Weakly Supervised Ordinal Support Vector Machine (WSO-SVM) to predict regional genetic alteration status within each GBM tumor using MRI. WSO-SVM was applied to a unique dataset of 318 image-localized biopsies with spatially matched multiparametric MRI from 74 GBM patients. The model was trained to predict the regional genetic alteration of three GBM driver genes (EGFR, PDGFRA, and PTEN) based on features extracted from the corresponding region of five MRI contrast images. For comparison, a variety of existing ML algorithms were also applied. The classification accuracy of each gene was compared between the different algorithms. The SHapley Additive exPlanations (SHAP) method was further applied to compute contribution scores of different contrast images. Finally, the trained WSO-SVM was used to generate prediction maps within the tumoral area of each patient to help visualize the intra-tumoral genetic heterogeneity. This study demonstrated the feasibility of using MRI and WSO-SVM to enable non-invasive prediction of intra-tumoral regional genetic alteration for each GBM patient, which can inform future adaptive therapies for individualized oncology."} {"id": "arxiv:2401.00133v1", "text": "In this paper, we present a signal processing framework for directed graphs. Unlike undirected graphs, a graph shift operator such as the adjacency matrix associated with a directed graph usually does not admit an orthogonal eigenbasis. This makes it challenging to define the Fourier transform. Our methodology leverages the polar decomposition to define two distinct eigendecompositions, each associated with different matrices derived from this decomposition. We propose to extend the frequency domain and introduce a Fourier transform that jointly encodes the spectral response of a signal for the two eigenbases from the polar decomposition. This allows us to define convolution following a standard routine. Our approach has two features: it is lossless as the shift operator can be fully recovered from factors of the polar decomposition. Moreover, it subsumes the traditional graph signal processing if the graph is directed. We present numerical results to show how the framework can be applied."} {"id": "arxiv:2401.00142v1", "text": "In recent work, Dao and Eisenbud define the notion of a Burch index, expanding the notion of Burch rings of Dao, Kobayashi, and Takahashi, and show that for any module over a ring of Burch index at least 2, its $n$th syzygy contains direct summands of the residue field for $n=4$ or $5$ and all $n\\geq 7$. We investigate how this behavior is explained by the bar resolution formed from appropriate differential graded (dg) resolutions, yielding a new proof that includes all $n\\geq 5$, which is sharp. When the module is Golod, we use instead the bar resolution formed from $A_\\infty$ resolutions to identify such $k$ summands explicitly for all $n\\geq 4$ and show that the number of these grows exponentially as the homological degree increases."} {"id": "arxiv:2401.00148v1", "text": "Autonomous vehicles increasingly utilize the vision-based perception module to acquire information about driving environments and detect obstacles. Correct detection and classification are important to ensure safe driving decisions. Existing works have demonstrated the feasibility of fooling the perception models such as object detectors and image classifiers with printed adversarial patches. However, most of them are indiscriminately offensive to every passing autonomous vehicle. In this paper, we propose TPatch, a physical adversarial patch triggered by acoustic signals. Unlike other adversarial patches, TPatch remains benign under normal circumstances but can be triggered to launch a hiding, creating or altering attack by a designed distortion introduced by signal injection attacks towards cameras. To avoid the suspicion of human drivers and make the attack practical and robust in the real world, we propose a content-based camouflage method and an attack robustness enhancement method to strengthen it. Evaluations with three object detectors, YOLO V3/V5 and Faster R-CNN, and eight image classifiers demonstrate the effectiveness of TPatch in both the simulation and the real world. We also discuss possible defenses at the sensor, algorithm, and system levels."} {"id": "arxiv:2401.00154v2", "text": "Recent studies emphasize that vehicular honking contributes to over 50% of noise pollution in developing urban and suburban areas. Frequent honking negatively impacts health, road safety, and the environment. Recognizing and classifying different vehicle honks could offer valuable insights into environmental noise pollution. Existing research on outdoor sound classification and honk detection lacks the ability to classify honks based on vehicle types, limiting contextual information inference for locations, areas, or traffic. Therefore, it becomes imperative to design a system that can detect and classify honks of different types of vehicles from which we can infer some contextual information. In this paper, we have developed a novel framework AClassiHonk that performs raw vehicular honk sensing, data labeling and classifies the honk into three major groups, i.e., light-weight vehicles, medium-weight vehicles, and heavy-weight vehicles. We collected the raw audio samples of different vehicular honking based on spatio-temporal characteristics and converted them into spectrogram images. We have proposed a deep learning-based Multi-label Autoencoder model (MAE) for automated labeling of the unlabeled data samples, which provides 97.64% accuracy in contrast to existing deep learning-based data labeling methods. Further, we have used various pre-trained models, namely Inception V3, ResNet50, MobileNet, ShuffleNet, and proposed an Ensembled Transfer Learning model (EnTL) for vehicle honks classification and performed comparative analysis. Results reveal that EnTL exhibits the best performance compared to pre-trained models and achieves 96.72% accuracy in our dataset. In addition, we have identified a context of a location based on these classified honk signatures in a city."} {"id": "arxiv:2401.00155v2", "text": "Occlusion presents a significant challenge in human pose estimation. The challenges posed by occlusion can be attributed to the following factors: 1) Data: The collection and annotation of occluded human pose samples are relatively challenging. 2) Feature: Occlusion can cause feature confusion due to the high similarity between the target person and interfering individuals. 3) Inference: Robust inference becomes challenging due to the loss of complete body structural information. The existing methods designed for occluded human pose estimation usually focus on addressing only one of these factors. In this paper, we propose a comprehensive framework DAG (Data, Attention, Graph) to address the performance degradation caused by occlusion. Specifically, we introduce the mask joints with instance paste data augmentation technique to simulate occlusion scenarios. Additionally, an Adaptive Discriminative Attention Module (ADAM) is proposed to effectively enhance the features of target individuals. Furthermore, we present the Feature-Guided Multi-Hop GCN (FGMP-GCN) to fully explore the prior knowledge of body structure and improve pose estimation results. Through extensive experiments conducted on three benchmark datasets for occluded human pose estimation, we demonstrate that the proposed method outperforms existing methods. Code and data will be publicly available."} {"id": "arxiv:2401.00129v5", "text": "We generalize Integration-By-Parts (IBP) and differential equations methods to de Sitter correlators related to inflation. While massive correlators in de Sitter spacetime are usually regarded as highly intricate, we find they have remarkably hidden concise structures from the perspective of IBP. We find the factorization of the IBP relations of each vertex integral family corresponding to $\\mathrm{d} τ_i$ integration. Furthermore, with a smart construction of master integrals, the universal formulas for iterative reduction and $\\mathrm{d} \\log$-form differential equations of arbitrary vertex integral family are presented and proved. These formulas dominate all tree-level de Sitter correlators and play a kernel role at the loop-level as well."} {"id": "arxiv:2401.00109v3", "text": "As time-series applications grow larger, there is increasing demand for symbolic representations that are compact, accurate, and scalable across many signals and computing resources. Current ABBA-based symbolic approximation methods produce high-quality, shape-preserving representations, but they handle each time series separately and sequentially. This means they do not ensure consistent symbols across different series and cannot fully exploit modern multicore systems and distributed-memory systems. This paper presents a joint symbolic time-series approximation method for large-scale time series. The proposed method decouples local compression from global digitization: (i) time series are partitioned into independent domains that can be compressed in parallel, and (ii) the resulting pieces are digitized using a shared global dictionary. To further improve scalability, we introduce a two-stage parallel digitization scheme, in which aggregation is first performed locally and then merged globally without requiring a full-data reassignment step. Extensive experiments on time-series datasets and large synthetic benchmarks show that our approach maintains competitive reconstruction quality while substantially reducing runtime. These results show that joint symbolic approximation can serve as an efficient, high-level parallel tool for analyzing large-scale temporal data."} {"id": "arxiv:2401.00173v1", "text": "We investigated the beat-to-beat fluctuation of the photoplethysmography (PPG) waveform. The motivation is that morphology variability extracted from the arterial blood pressure (ABP) has been found to correlate with baseline condition and short-term surgical outcome of the patients undergoing liver transplant surgery. Numerous interactions of physiological mechanisms regulating the cardiovascular system could underlie the variability of morphology. We used the unsupervised manifold learning algorithm, Dynamic Diffusion Map, to quantify the multivariate waveform morphological variation. Due to the physical principle of light absorption, PPG waveform signals are more susceptible to artifact and are nominally used only for visual inspection of data quality in clinical medical environment. But on the other hand, the noninvasive, easy-to-use nature of PPG grants a wider range of biomedical application, which inspired us to investigate the variability of morphology information from PPG waveform signal. We developed data analysis techniques to improve the performance and validated with the real-life clinical database."} {"id": "arxiv:2401.00136v1", "text": "We extend prior work to derive three additional M-1-dimensional integral representations--over the interval $[0,1]$ --for products of M Slater orbitals that allows their magnitudes of coordinate vector differences (square roots of polynomials) $|{\\bf x}_{1}-{\\bf x}_{2}|=\\sqrt{x_{1}^{2}-2x_{1}x_{2}\\cosθ+x_{2}^{2}}$ to be moved from disjoint products of functions into a single quadratic form whose square my be completed. This provides more alternatives to Fourier transforms that introduce a 3M-dimensional momentum integral for those products of Slater orbitals, followed by another set of M-1-dimensional integral representations to combine those denominators into one denominator having a single (momentum) quadratic form. The current work is also slightly more compact than Gaussian transforms that introduce an M-dimensional integral for products of M Slater orbitals. We have found that two of these M-1-dimensional integral representations over the interval $[0,1]$ are numerically stable, as was the prior version having integrals running over the interval $[0,\\infty]$, and one does not need to test for a sufficiently large upper integration limit. For analytical reductions of integrals arising from any of the three, however, there is the possible drawback for large M of there being fewer tabled integrals over $[0,1]$ than over $[0,\\infty]$. These representations have integration variables within square roots as arguments of Macdonald functions. In a number of cases, these may be converted to Meijer G-functions for which a single tabled integral exists over the interval $[0,\\infty]$ of the prior paper, and from which other forms may be found. Finally, we introduce a fourth integral representation that is not easily generalizable to large M, but may well provide a bridge for finding the requisite integrals for such Meijer G-functions over $[0,1]$."} {"id": "arxiv:2401.00161v1", "text": "The hybrid neural differentiable models mark a significant advancement in the field of scientific machine learning. These models, integrating numerical representations of known physics into deep neural networks, offer enhanced predictive capabilities and show great potential for data-driven modeling of complex physical systems. However, a critical and yet unaddressed challenge lies in the quantification of inherent uncertainties stemming from multiple sources. Addressing this gap, we introduce a novel method, DiffHybrid-UQ, for effective and efficient uncertainty propagation and estimation in hybrid neural differentiable models, leveraging the strengths of deep ensemble Bayesian learning and nonlinear transformations. Specifically, our approach effectively discerns and quantifies both aleatoric uncertainties, arising from data noise, and epistemic uncertainties, resulting from model-form discrepancies and data sparsity. This is achieved within a Bayesian model averaging framework, where aleatoric uncertainties are modeled through hybrid neural models. The unscented transformation plays a pivotal role in enabling the flow of these uncertainties through the nonlinear functions within the hybrid model. In contrast, epistemic uncertainties are estimated using an ensemble of stochastic gradient descent (SGD) trajectories. This approach offers a practical approximation to the posterior distribution of both the network parameters and the physical parameters. Notably, the DiffHybrid-UQ framework is designed for simplicity in implementation and high scalability, making it suitable for parallel computing environments. The merits of the proposed method have been demonstrated through problems governed by both ordinary and partial differentiable equations."} {"id": "arxiv:2401.00165v2", "text": "In open-domain Question Answering (QA), dense retrieval is crucial for finding relevant passages for answer generation. Typically, contrastive learning is used to train a retrieval model that maps passages and queries to the same semantic space. The objective is to make similar ones closer and dissimilar ones further apart. However, training such a system is challenging due to the false negative issue, where relevant passages may be missed during data annotation. Hard negative sampling, which is commonly used to improve contrastive learning, can introduce more noise in training. This is because hard negatives are those closer to a given query, and thus more likely to be false negatives. To address this issue, we propose a novel contrastive confidence regularizer for Noise Contrastive Estimation (NCE) loss, a commonly used loss for dense retrieval. Our analysis shows that the regularizer helps dense retrieval models be more robust against false negatives with a theoretical guarantee. Additionally, we propose a model-agnostic method to filter out noisy negative passages in the dataset, improving any downstream dense retrieval models. Through experiments on three datasets, we demonstrate that our method achieves better retrieval performance in comparison to existing state-of-the-art dense retrieval systems."} {"id": "arxiv:2401.00168v1", "text": "In this paper, we scale evolutionary algorithms to high-dimensional optimization problems that deceptively possess a low effective dimensionality (certain dimensions do not significantly affect the objective function). To this end, an instantiation of the multiform optimization paradigm is presented, where multiple low-dimensional counterparts of a target high-dimensional task are generated via random embeddings. Since the exact relationship between the auxiliary (low-dimensional) tasks and the target is a priori unknown, a multiform evolutionary algorithm is developed for unifying all formulations into a single multi-task setting. The resultant joint optimization enables the target task to efficiently reuse solutions evolved across various low-dimensional searches via cross-form genetic transfers, hence speeding up overall convergence characteristics. To validate the overall efficacy of our proposed algorithmic framework, comprehensive experimental studies are carried out on well-known continuous benchmark functions as well as a set of practical problems in the hyper-parameter tuning of machine learning models and deep learning models in classification tasks and Predator-Prey games, respectively."} {"id": "arxiv:2401.00182v1", "text": "For a finite valued field extension $(L/K,v)$ we describe the problem of find sets of generators for the corresponding extension $\\mathcal O_L/\\mathcal O_K$ of valuation rings. The main tool to obtain such sets are complete sets of (key) polynomials. We show that when the initial index coincide with the ramification index, sequences of key polynomials naturally give rise to sets of generators. We use this to prove Knaf's conjecture for pure extensions."} {"id": "arxiv:2401.00185v1", "text": "Utilizing observations from the New Vacuum Solar Telescope (NVST), Solar Dynamics Observatory (SDO), and Solar Terrestrial Relations Observatory-Ahead (STEREO-A), we investigate the event from two distinct observational perspectives: on the solar disk using NVST and SDO, and on the solar limb using STEREO-A. We employ both a non-linear force-free field model and a potential field model to reconstruct the coronal magnetic field, aiming to understand its magnetic properties. Two precursor jet-like activities were observed before the eruption, displaying an untwisted rotation. The second activity released an estimated twist of over two turns. During these two jet-like activities, Y-shaped brightenings, newly emerging magnetic flux accompanied by magnetic cancellation, and the formation of newly moving fibrils were identified. Combining these observational features, it can be inferred that these two precursor jet-like activities released the magnetic field constraining the filament and were triggered by newly emerging magnetic flux. Before the filament eruption, it was observed that some moving flows had been ejected from the site as the onset of two jet-like activities, indicating the same physical process as two jet-like activities. Extrapolations revealed that the filament laid under the height of the decay index of 1.0 and had strong magnetic field (540 Gauss) and a high twisted number (2.4 turns) before the eruption. An apparent rotational motion was observed during the filament eruption. We deduce that the solar filament, exhibiting an inverted U-shape, is a significantly twisted flux rope. The eruption of the filament was initiated by the release of constraining magnetic fields through continuous magnetic reconnection. This reconnection process was triggered by the emergence of newly magnetic flux."} {"id": "arxiv:2401.00189v1", "text": "In this paper, we introduce semi-infinite tensor complementarity problem to provide an approach for considering a more realistic situation of the problem. We prove the necessary and sufficient conditions for the existence of the solution set. In this context, we study the error bounds of the solution set in terms of residual function."} {"id": "arxiv:2401.00190v1", "text": "The FIR distribution at high Galactic latitudes, observed with Planck, is filamentary with coherent structures in polarization. These structures are also closely related to HI filaments with coherent velocity structures. There is a long-standing debate about the physical nature of these structures. They are considered either as velocity caustics, fluctuations engraved by the turbulent velocity field or as cold three-dimensional density structures in the interstellar medium (ISM). We discuss different approaches to data analysis and interpretation in order to work out the differences. We considered mathematical preliminaries for the derivation of caustics that characterize filamentary structures in the ISM. Using the Hessian operator, we traced individual FIR filamentary structures in HI from channel maps as observed and alternatively from data that are provided by the velocity decomposition algorithm (VDA). VDA is claimed to separate velocity caustics from density effects. Based on the strict mathematical definition, the so-called velocity caustics are not actually caustics. These VDA data products may contain caustics in the same way as the original HI observations. Caustics derived by a Hessian analysis of both databases are nearly identical with a correlation coefficient of 98%. However, the VDA algorithm leads to a 30% increase in the alignment uncertainties when fitting FIR/HI orientation angles. We used HI absorption data to constrain the physical nature of FIR/HI filaments and determine spin temperatures and volume densities of FIR/HI filaments. HI filaments exist as CNM structures; outside the filaments no CNM absorption is detectable. The CNM in the diffuse ISM is exclusively located in filaments with FIR counterparts. These filaments at high Galactic latitudes exist as cold density structures; velocity crowding effects are negligible."} {"id": "arxiv:2401.00195v1", "text": "Many problems can be solved in two ways: either by adapting an existing solution, or by exchanging it for a new one. To investigate under what conditions people consider new solutions, we traced their information acquisition processes in a simulated mechanical engineering task. Within a multi-step optimisation procedure, participants could either adapt the properties of a currently used machine component, or exchange this component for a new one. They had the opportunity to check whether the solutions met a set of requirements, which was varied systematically. We investigated whether participants would consistently check both solutions, or whether they would satisfice, ignoring the new solution as long as the current one was good enough. The results clearly refuted consistent checking, but only partly confirmed satisficing. On the one hand, participants indeed checked the new solution least often when the current one was applicable without problems. On the other hand, in this case the new solution still was not fully ignored. However, the latter finding could be traced back to a few participants who diverged from our anticipated strategy of first checking the current solution, and directly went for the new one. The results suggest that in Adapt/Exchange decisions, people do not usually check both solutions in an unbiased manner, but rely on existing solutions as long as they are good enough."} {"id": "arxiv:2401.00111v1", "text": "Here, we review some quantum architectures designed for the engineering of the N00N state, a bipartite maximally entangled state crucial in quantum metrology applications. The fundamental concept underlying these schemes is the transformation of the initial state $|N\\rangle \\otimes |0\\rangle$ to the N00N state $\\frac{1}{\\sqrt{2}} (|N\\rangle \\otimes|0\\rangle +|0\\rangle \\otimes|N\\rangle)$, where $|N\\rangle$ and $|0\\rangle$ are the Fock states with $N$ and $0$ excitations. We show that this state can be generated as a superposition of modes of quantum light, a combination of light and motion, or a superposition of two spin ensembles. The approach discussed here can generate mesoscopic and macroscopic entangled states, such as entangled coherent and squeezed states, as well. We show that a large class of maximally entangled states can be achieved in such an architecture. The extension of these state engineering methods to the multi-mode setting is also discussed."} {"id": "arxiv:2401.00113v1", "text": "Late-type stars are the most abundant in the galactic stellar population. These stars, with a similar internal structure to the Sun, are expected to have solar-like atmospheres. Investigating the stellar parameters and chemical abundances on late-type stars is essential to provide valuable constraints about stellar age, chemical evolution, and atmosphere of exoplanets. In this work, we present the study of the Near-UV and optical spectroscopic observation of three late-type stars: HR 8038, AC Her, and HD 76446, as obtained from the 36-inch MIRA/Oliver Observing Station. We derived surface temperature, gravity, metallicity, and the chemical abundances of light element Carbon in the stellar atmosphere. The elemental abundance of the Carbon for HR 8038, AC Her, and HD 76446 are derived to be 95%, 97%, and 108%, respectively, of the solar value."} {"id": "arxiv:2401.00123v2", "text": "This paper is an extended and reworked version of a short course given by the author at ''Uzbekistan-Ukrainian readings in stochastic processes'', Tashkent-Kyiv, 2022, and was prepared for a special issue of ''Theory of stochastic processes'', devoted to publishing lecture notes from the aforementioned workshop. The survey is devoted to operator splitting methods in the abstract formulation and their applications in probability. While the survey is focused on multiplicative methods, the BCH formula is used to discuss exponential splitting methods and a short informal introduction to additive splitting is presented. We introduce frameworks and available deterministic and probabilistic results and concentrate on constructing a wide picture of the field of operator splitting methods, providing a rigorous description in the setting of abstract Cauchy problems and an informal discussion for further and parallel advances. Some limitations and common difficulties are listed, as well as examples of works that provide solutions or hints. No new results are given. The bibliography contains illustrative deterministic examples and a selection of probability-related works."} {"id": "arxiv:2401.00151v1", "text": "The proliferation of images captured from millions of cameras and the advancement of facial recognition (FR) technology have made the abuse of FR a severe privacy threat. Existing works typically rely on obfuscation, synthesis, or adversarial examples to modify faces in images to achieve anti-facial recognition (AFR). However, the unmodified images captured by camera modules that contain sensitive personally identifiable information (PII) could still be leaked. In this paper, we propose a novel approach, CamPro, to capture inborn AFR images. CamPro enables well-packed commodity camera modules to produce images that contain little PII and yet still contain enough information to support other non-sensitive vision applications, such as person detection. Specifically, CamPro tunes the configuration setup inside the camera image signal processor (ISP), i.e., color correction matrix and gamma correction, to achieve AFR, and designs an image enhancer to keep the image quality for possible human viewers. We implemented and validated CamPro on a proof-of-concept camera, and our experiments demonstrate its effectiveness on ten state-of-the-art black-box FR models. The results show that CamPro images can significantly reduce face identification accuracy to 0.3\\% while having little impact on the targeted non-sensitive vision application. Furthermore, we find that CamPro is resilient to adaptive attackers who have re-trained their FR models using images generated by CamPro, even with full knowledge of privacy-preserving ISP parameters."} {"id": "arxiv:2401.00159v1", "text": "Progression of hip osteoarthritis (hip OA) leads to pain and disability, likely leading to surgical treatment such as hip arthroplasty at the terminal stage. The severity of hip OA is often classified using the Crowe and Kellgren-Lawrence (KL) classifications. However, as the classification is subjective, we aimed to develop an automated approach to classify the disease severity based on the two grades using digitally-reconstructed radiographs (DRRs) from CT images. Automatic grading of the hip OA severity was performed using deep learning-based models. The models were trained to predict the disease grade using two grading schemes, i.e., predicting the Crowe and KL grades separately, and predicting a new ordinal label combining both grades and representing the disease progression of hip OA. The models were trained in classification and regression settings. In addition, the model uncertainty was estimated and validated as a predictor of classification accuracy. The models were trained and validated on a database of 197 hip OA patients, and externally validated on 52 patients. The model accuracy was evaluated using exact class accuracy (ECA), one-neighbor class accuracy (ONCA), and balanced accuracy.The deep learning models produced a comparable accuracy of approximately 0.65 (ECA) and 0.95 (ONCA) in the classification and regression settings. The model uncertainty was significantly larger in cases with large classification errors (P<6e-3). In this study, an automatic approach for grading hip OA severity from CT images was developed. The models have shown comparable performance with high ONCA, which facilitates automated grading in large-scale CT databases and indicates the potential for further disease progression analysis. Classification accuracy was correlated with the model uncertainty, which would allow for the prediction of classification errors."} {"id": "arxiv:2401.00198v1", "text": "In this paper, we study the stability of traveling wave solutions arising from a credit rating migration problem with a free boundary, After some transformations, we turn the Free Boundary Problem into a fully nonlinear parabolic problem on a fixed domain and establish a rigorous stability analysis of the equilibrium in an exponentially weighted function space. It implies the convergence of the discounted value of bonds that stands as an attenuated traveling wave solution."} {"id": "arxiv:2401.00108v2", "text": "In this work, we consider constrained stochastic optimization problems under hidden convexity, i.e., those that admit a convex reformulation via non-linear (but invertible) map $c(\\cdot)$. A number of non-convex problems ranging from optimal control, revenue and inventory management, to convex reinforcement learning all admit such a hidden convex structure. Unfortunately, in the majority of applications considered, the map $c(\\cdot)$ is unavailable or implicit; therefore, directly solving the convex reformulation is not possible. On the other hand, the stochastic gradients with respect to the original variable are often easy to obtain. Motivated by these observations, we examine the basic projected stochastic (sub-) gradient methods for solving such problems under hidden convexity. We provide the first sample complexity guarantees for global convergence in smooth and non-smooth settings. Additionally, in the smooth setting, we improve our results to the last iterate convergence in terms of function value gap using the momentum variant of projected stochastic gradient descent."} {"id": "arxiv:2401.00162v3", "text": "The sparsity of reward feedback remains a challenging problem in online deep reinforcement learning (DRL). Previous approaches have utilized offline demonstrations to achieve impressive results in multiple hard tasks. However, these approaches place high demands on demonstration quality, and obtaining expert-like actions is often costly and unrealistic. To tackle these problems, we propose a simple and efficient algorithm called Policy Optimization with Smooth Guidance (POSG), which leverages a small set of state-only demonstrations (where expert action information is not included in demonstrations) to indirectly make approximate and feasible long-term credit assignments and facilitate exploration. Specifically, we first design a trajectory-importance evaluation mechanism to determine the quality of the current trajectory against demonstrations. Then, we introduce a guidance reward computation technology based on trajectory importance to measure the impact of each state-action pair, fusing the demonstrator's state distribution with reward information into the guidance reward. We theoretically analyze the performance improvement caused by smooth guidance rewards and derive a new worst-case lower bound on the performance improvement. Extensive results demonstrate POSG's significant advantages in control performance and convergence speed in four sparse-reward environments, including the grid-world maze, Hopper-v4, HalfCheetah-v4, and Ant maze. Notably, the specific metrics and quantifiable results are investigated to demonstrate the superiority of POSG."} {"id": "arxiv:2401.00179v1", "text": "Gas-particle flows are commonly simulated through two-fluid model at industrial-scale. However, these simulations need very fine grid to have accurate flow predictions, which is prohibitively demanding in terms of computational resources. To circumvent this problem, the filtered two-fluid model has been developed, where large-scale flow field is numerically resolved and small-scale fluctuations are accounted for through subgrid-scale modeling. In this study, we have performed fine-grid two-fluid simulations of dilute gas-particle flows in periodic domains and applied explicit filtering to generate datasets. Then, these datasets have been used to develop artificial neural network (ANN) models for closures such as the filtered drag force and solid phase stress for the filtered two-fluid model. The set of input variables for the subgrid drag force ANN model that has been found previously to work well for dense flow regimes is found to work as well for the dilute regime. In addition, we present a Galilean invariant tensor basis neural network (TBNN) model for the filtered solid phase stress which can capture nicely the anisotropic nature of the solid phase stress arising from subgrid-scale velocity fluctuations. Finally, the predictions provided by this new TBNN model are compared with those obtained from a simple eddy-viscosity ANN model."} {"id": "arxiv:2401.00127v1", "text": "$ $The synergy of language and vision models has given rise to Large Language and Vision Assistant models (LLVAs), designed to engage users in rich conversational experiences intertwined with image-based queries. These comprehensive multimodal models seamlessly integrate vision encoders with Large Language Models (LLMs), expanding their applications in general-purpose language and visual comprehension. The advent of Large Multimodal Models (LMMs) heralds a new era in Artificial Intelligence (AI) assistance, extending the horizons of AI utilization. This paper takes a unique perspective on LMMs, exploring their efficacy in performing image classification tasks using tailored prompts designed for specific datasets. We also investigate the LLVAs zero-shot learning capabilities. Our study includes a benchmarking analysis across four diverse datasets: MNIST, Cats Vs. Dogs, Hymnoptera (Ants Vs. Bees), and an unconventional dataset comprising Pox Vs. Non-Pox skin images. The results of our experiments demonstrate the model's remarkable performance, achieving classification accuracies of 85\\%, 100\\%, 77\\%, and 79\\% for the respective datasets without any fine-tuning. To bolster our analysis, we assess the model's performance post fine-tuning for specific tasks. In one instance, fine-tuning is conducted over a dataset comprising images of faces of children with and without autism. Prior to fine-tuning, the model demonstrated a test accuracy of 55\\%, which significantly improved to 83\\% post fine-tuning. These results, coupled with our prior findings, underscore the transformative potential of LLVAs and their versatile applications in real-world scenarios."} {"id": "arxiv:2401.00164v2", "text": "We study solutions to systems of stream inclusions of the form 'f in T(f)', where the nondeterministic transformer 'T' on omega-infinite streams is assumed to be causal in the sense that elements in output streams are determined by a finite prefix of inputs. We first establish a correspondence between logic-based causality and metric-based contraction. Based on this causality-contraction connection we then apply fixpoint principles to the spherically complete ultrametric space of streams to construct solutions of stream inclusions. The underlying fixpoint iterations induce fixpoint induction principles to reason about these solutions.In addition, the fixpoint approximation provides an anytime algorithm with which finite prefixes of solutions can be calculated. These developments are illustrated for some central concepts of system design."} {"id": "arxiv:2401.00130v2", "text": "Recently, Amnon Neeman settled a bold conjecture by Antieau, Gepner, and Heller regarding the relationship between the regularity of finite-dimensional noetherian schemes and the existence of bounded $t$-structures on their derived categories of perfect complexes. In this paper, using different methods, we prove some very general results about the existence of bounded $t$-structures on (not necessarily algebraic or topological) triangulated categories and their invariance under completion. We show that if the opposite category of an essentially small triangulated category has finite finitistic dimension in our sense, then the existence of a bounded t-structure on it forces it to be equal to its completion. We also prove a parallel result regarding the equivalence of all bounded t-structures on any intermediate triangulated category between the starting category and its completion. Our general treatment, when specialized to the case of schemes, immediately gives us Neeman's theorem as an application and significantly generalizes another remarkable theorem by Neeman about the equivalence of bounded $t$-structures on the bounded derived categories of coherent sheaves. When specialized to other cases like associative rings, nonpositive DG-rings, connective $\\mathbb{E}_1$-rings, triangulated categories without models, etc., we get many other applications. Under mild finiteness assumptions, these results not only give a categorical obstruction (the singularity category in our sense) to the existence of bounded $t$-structures on a triangulated category, but also provide plenty of triangulated categories on which all bounded $t$-structures are equivalent. The strategy used in our treatment is introducing a new concept of finitistic dimension for triangulated categories and lifting $t$-structures along completions of triangulated categories."} {"id": "arxiv:2401.00146v1", "text": "Quantum cryptography is now considered as a promising technology due to its promise of unconditional security. In recent years, rigorous work is being done for the experimental realization of quantum key distribution (QKD) protocols to realize secure networks. Among various QKD protocols, coherent one way and differential phase shift QKD protocols have undergone rapid experimental developments due to the ease of experimental implementations with the present available technology. In this work, we have experimentally realized optical fiber based coherent one way and differential phase shift QKD protocols at telecom wavelength. Both protocols belong to a class of protocols named as distributed phase reference protocol in which weak coherent pulses are used to encode the information. Further, we have analyzed the key rates with respect to different parameters such distance, disclose rate, compression ratio and detector dead time."} {"id": "arxiv:2401.00169v3", "text": "Wiener spaces are in many ways the decisive setting for fundamental results on Gaussian measures: large deviations (Schilder), quasi-invariance (Cameron--Martin), differential calculus (Malliavin), support description (Stroock--Varadhan), concentration of measure (Fernique), etc. Analogues of these classical results have been derived in the \"enhanced\" context of Gaussian rough paths and, more recently, regularity structures equipped with Gaussian models. The aim of this article is to propose a similar notion directly on this enhanced level - an abstract Wiener model space - that encompasses the aforementioned. More specifically, we focus here on enhanced Schilder type results, Cameron--Martin shifts and Fernique estimates, offering a somewhat unified view on results of Friz--Victoir and Hairer--Weber."} {"id": "arxiv:2401.00121v2", "text": "We propose a contour integral-based algorithm for computing a few singular values of a matrix or a few generalized singular values of a matrix pair. Mathematically, the generalized singular values of a matrix pair are the eigenvalues of an equivalent Hermitian-definite matrix pencil, known as the Jordan-Wielandt matrix pencil. However, direct application of the FEAST algorithm does not fully exploit the structure of this problem. We analyze several projection strategies on the Jordan-Wielandt matrix pencil, and propose an effective and robust scheme tailored to GSVD. Both theoretical analysis and numerical experiments demonstrate that our algorithm achieves rapid convergence and satisfactory accuracy."} {"id": "arxiv:2401.00144v2", "text": "We investigate the tachyonic instability of Kerr-Newman (KN) black hole with a rotation parameter $a$ in the Einstein-Chern-Simons-scalar theory coupled with a quadratic massive scalar field. This instability analysis corresponds to exploring the onset of spontaneous scalarization for KN black holes. First, we find no $a$-bound for $α<0$ case by considering (1+1)-dimensional analytical method. A direct numerical method is adopted to explore (2+1)-dimensional time evolution of a massive scalar perturbation with positive and negative $α$ to obtain threshold curves numerically. We obtain threshold curves $α_{\\rm th}(a)$ of tachyonic instability for positive $α$ without any $a$-bounds. We expect to find the same threshold curves $α_{\\rm th}(a)$ of tachyonic instability for negative $α$ without any $a$-bound because its linearized scalar theory is invariant under the transformation of $α\\to -α$ and $θ\\to -θ$. In addition, it is found that the scalar mass term suppresses tachyonic instability of KN black holes."} {"id": "arxiv:2401.00131v3", "text": "In this article, we investigate periodically driven open quantum systems within the framework of Floquet-Lindblad master equations. Specifically, we discuss Lindblad master equations in the presence of a coherent, time-periodic driving and establish their general spectral features. We also clarify the notions of transient and non-decaying solutions from this spectral perspective, and then prove that any physical system described by a Floquet-Lindblad equation must have at least one \\textit{physical} non-equilibrium steady state (NESS), corresponding to an eigenoperator of the Floquet-Lindblad evolution superoperator $\\mathcal{U}_F$ with unit eigenvalue. Since the Floquet-Lindblad formalism encapsulates the entire information regarding the NESS, it in principle enables us to obtain non-linear effects to all orders at once. The Floquet-Lindblad formalism thus provides a powerful tool for studying driven-dissipative solid-state systems, which we illustrate by deriving the nonlinear optical response of a simple two-band model of an insulating solid and comparing it with prior results established through Keldysh techniques."} {"id": "arxiv:2401.00163v1", "text": "Backdoor attacks in the traditional graph neural networks (GNNs) field are easily detectable due to the dilemma of confusing labels. To explore the backdoor vulnerability of GNNs and create a more stealthy backdoor attack method, a clean-label graph backdoor attack method(CGBA) in the node classification task is proposed in this paper. Differently from existing backdoor attack methods, CGBA requires neither modification of node labels nor graph structure. Specifically, to solve the problem of inconsistency between the contents and labels of the samples, CGBA selects poisoning samples in a specific target class and uses the label of sample as the target label (i.e., clean-label) after injecting triggers into the target samples. To guarantee the similarity of neighboring nodes, the raw features of the nodes are elaborately picked as triggers to further improve the concealment of the triggers. Extensive experiments results show the effectiveness of our method. When the poisoning rate is 0.04, CGBA can achieve an average attack success rate of 87.8%, 98.9%, 89.1%, and 98.5%, respectively."} {"id": "arxiv:2401.00180v1", "text": "This paper proposes a cyber-resilient distributed control strategy equipped with attack detection capabilities for islanded AC microgrids in the presence of bounded stealthy cyber attacks affecting both frequency and power information exchanged among neighboring distributed generators (DGs). The proposed control methodology relies on the construction of an auxiliary layer and the establishment of effective inter-layer cooperation between the actual DGs in the control layer and the virtual DGs in the auxiliary layer. This cooperation aims to achieve robust frequency restoration and proportional active power-sharing. It is shown that the in situ presence of a concealed auxiliary layer not only guarantees resilience against stealthy bounded attacks on both frequency and power-sharing but also facilitates a network-enabled attack identification mechanism. The paper provides rigorous proof of the stability of the closed-loop system and derives bounds for frequency and power deviations under attack conditions, offering insights into the impact of the attack signal, control and pinning gains, and network connectivity on the system's convergence properties. The performance of the proposed controllers is illustrated by simulating a networked islanded AC microgrid in a Simulink environment showcasing both attributes of attack resilience and attack detection."} {"id": "arxiv:2401.00188v1", "text": "We propose a discrete-time econometric model that combines autoregressive filters with factor regressions to predict stock returns for portfolio optimisation purposes. In particular, we test both robust linear regressions and general additive models on two different investment universes composed of the Dow Jones Industrial Average and the Standard & Poor's 500 indexes, and we compare the out-of-sample performances of mean-CVaR optimal portfolios over a horizon of six years. The results show a substantial improvement in portfolio performances when the factor model is estimated with general additive models."} {"id": "arxiv:2401.00192v1", "text": "In this letter, we propose a deep-unfolding-based framework (DUNet) to maximize the secrecy rate in reconfigurable intelligent surface (RIS) empowered multi-user wireless networks. To tailor DUNet, first we relax the problem, decouple it into beamforming and phase shift subproblems, and propose an alternative optimization (AO) based solution for the relaxed problem. Second, we apply Karush-Kuhn-Tucker (KKT) conditions to obtain a closed-form solutions for the beamforming and the phase shift. Using deep-unfolding mechanism, we transform the closed-form solutions into a deep learning model (i.e., DUNet) that achieves a comparable performance to that of AO in terms of accuracy and about 25.6 times faster."} {"id": "arxiv:2401.00200v1", "text": "We present a framework to assist therapists and children with autism spectrum disorder in their Applied Behavioral Analysis (ABA) therapy. The framework was designed in collaboration with Spazio Autismo, an autism center in Mantova, Italy. The framework is a first step toward transitioning from the current paper-based to fully digital-supported therapy. We evaluated the framework over four months with 18 children diagnosed with classic autism, ranging from 4 to 7 years old. The framework integrates a mobile app that children and therapists use during the sessions with a backend for managing therapy workflow and monitoring progress. Our preliminary results show that the framework can improve the efficacy of the therapy sessions, reducing non-therapeutic time, increasing patient focus, and quickening the completion of the assigned objectives. It can also support therapists in preparing learning materials, data acquisition, and reporting. Finally, the framework demonstrated improved privacy and security of patients' data while maintaining reliability."} {"id": "arxiv:2401.00105v2", "text": "In this paper, we elaborate on correctly predicting Échelle spectrograms by employing the fully three-dimensional representation of Snell's law to model the effects of prisms as cross-dispersers in Échelle spectrographs. We find that it is not sufficient to simply apply the frequently used trigonometric prism dispersion equation to describe recorded spectra. This vector equation approach is not limited to a single dispersive element when modelling multi-prism cross-disperser configurations. Our results help to understand the main levers in an Échelle spectrograph as well as contribute to auto-calibration algorithms for minimizing calibration efforts in daily operation."} {"id": "arxiv:2401.00124v2", "text": "Generative artificial intelligence (GAI) has emerged as a rapidly burgeoning field demonstrating significant potential in creating diverse contents intelligently and automatically. To support such artificial intelligence-generated content (AIGC) services, future communication systems should fulfill much more stringent requirements (including data rate, throughput, latency, etc.) with limited yet precious spectrum resources. To tackle this challenge, semantic communication (SemCom), dramatically reducing resource consumption via extracting and transmitting semantics, has been deemed as a revolutionary communication scheme. The advanced GAI algorithms facilitate SemCom on sophisticated intelligence for model training, knowledge base construction and channel adaption. Furthermore, GAI algorithms also play an important role in the management of SemCom networks. In this survey, we first overview the basics of GAI and SemCom as well as the synergies of the two technologies. Especially, the GAI-driven SemCom framework is presented, where many GAI models for information creation, SemCom-enabled information transmission and information effectiveness for AIGC are discussed separately. We then delve into the GAI-driven SemCom network management involving with novel management layers, knowledge management, and resource allocation. Finally, we envision several promising use cases, i.e., autonomous driving, smart city, and the Metaverse for a more comprehensive exploration."} {"id": "arxiv:2401.00125v1", "text": "Although planning is a crucial component of the autonomous driving stack, researchers have yet to develop robust planning algorithms that are capable of safely handling the diverse range of possible driving scenarios. Learning-based planners suffer from overfitting and poor long-tail performance. On the other hand, rule-based planners generalize well, but might fail to handle scenarios that require complex driving maneuvers. To address these limitations, we investigate the possibility of leveraging the common-sense reasoning capabilities of Large Language Models (LLMs) such as GPT4 and Llama2 to generate plans for self-driving vehicles. In particular, we develop a novel hybrid planner that leverages a conventional rule-based planner in conjunction with an LLM-based planner. Guided by commonsense reasoning abilities of LLMs, our approach navigates complex scenarios which existing planners struggle with, produces well-reasoned outputs while also remaining grounded through working alongside the rule-based approach. Through extensive evaluation on the nuPlan benchmark, we achieve state-of-the-art performance, outperforming all existing pure learning- and rule-based methods across most metrics. Our code will be available at https://llmassist.github.io."} {"id": "arxiv:2401.00145v1", "text": "Miniaturized two-dimensional scanning mirror based on microelectromechanical systems (MEMS) technology has great potential in automotive industry, consumer electronics, and biomedicine, etc. Due to its high frequency and large angle, resonant scanning is the mainstream in all MEMS actuation mechanisms, such as harmonic resonant electromagnetic scanner and parametric resonant electrostatic scanner. Although electrostatic scanner has the advantages of low power consumption and IC process compatibility, some shortcomings of parametric resonance, including double frequency of driver electronics and additional feedback control or frequency stabilization system, limit its further application. The symmetry of coplanar electrostatic comb actuator is broken in this paper, and harmonic resonant electrostatic scanner with excellent performance is realized. Further, through adopting mechanical filter, two-dimensional scanning can be achieved through one set of actuators, which avoids the problem that two sets (each for one dimension) of electrostatic actuators must be insulated each other through complicated and expensive processes. A two-dimensional MEMS scanner based on symmetry breaking and mechanical filter was proposed and demonstrated. Multiple scanning modes can be achieved through selective control of a set of four identical actuators."} {"id": "arxiv:2401.00157v2", "text": "Metastability in open system dynamics describes the phenomena of initial relaxation to longlived metastable states before decaying to the asymptotic stable states. It has been predicted in continuous-time stochastic dynamics of both classical and quantum systems. Here we present a general theory of metastability in discrete-time open quantum dynamics, described by sequential quantum channels. We focus on a general class of quantum channels on a target system, induced by an ancilla system with a pure-dephasing coupling to the target system and under Ramsey sequences. Interesting metastable behaviors are predicted and numerically demonstrated by decomposing the average dynamics into stochastic trajectories. Examples and applications are also discussed."} {"id": "arxiv:2401.00194v2", "text": "Modulo sampling (MS) has been recently introduced to enhance the dynamic range of conventional ADCs by applying a modulo operator before sampling. This paper examines the identifiability of a measurement model where measurements are taken using a discrete Fourier transform (DFT) sensing matrix, followed by a modulo operator (modulo-DFT). Firstly, we derive a necessary and sufficient condition for the unique identification of the modulo-DFT sensing model based on the number of measurements and the indices of zero elements in the original signal. Then, we conduct a deeper analysis of three specific cases: when the number of measurements is a power of $2$, a prime number, and twice a prime number. Additionally, we investigate the identifiability of periodic bandlimited (PBL) signals under MS, which can be considered as the modulo-DFT sensing model with additional symmetric and conjugate constraints on the original signal. We also provide a necessary and sufficient condition based solely on the number of samples in one period for the unique identification of the PBL signal under MS, though with an ambiguity in the direct current (DC) component. Furthermore, we show that when the oversampling factor exceeds $3(1+1/P)$, the PBL signal can be uniquely identified with an ambiguity in the DC component, where $P$ is the number of harmonics, including the fundamental component, in the positive frequency part. Finally, we also present a recovery algorithm that estimates the original signal by solving integer linear equations, and we conduct simulations to validate our conclusions."} {"id": "arxiv:2401.00138v2", "text": "The antiferromagnetic layered compound EuCd$_2$As$_2$ is widely considered as a leading candidate of ideal Weyl semimetal, featuring a single pair of Weyl nodes in its field-induced ferromagnetic (FM) state. Nevertheless, this view has recently been challenged by an optical spectroscopy study, which suggests that it is a magnetic semiconductor. In this study, we have successfully synthesized highly insulating EuCd$_2$As$_2$ crystals with carrier density reaching as low as $2\\times 10^{15}$ $\\text{cm}^{-3}$. The magneto-transport measurements revealed a progressive decrease of the anomalous Hall conductivity (AHC) by several orders of magnitude as the carrier density decreases. This behavior contradicts with what is expected from the intrinsic AHC generated by the Weyl points, which is independent of carrier density as the Fermi level approaches the charge neutrality point. In contrast, the scaling relationship between AHC and longitudinal conductivity aligns with the characteristics of variable range hopping insulators. Our results suggest that EuCd$_2$As$_2$ is a magnetic semiconductor rather than a topological Weyl semimetal."} {"id": "arxiv:2401.00134v1", "text": "Training large-scale language models is increasingly critical in various domains, but it is hindered by frequent failures, leading to significant time and economic costs. Current failure recovery methods in cloud-based settings inadequately address the diverse and complex scenarios that arise, focusing narrowly on erasing downtime for individual tasks without considering the overall cost impact on a cluster. We introduce Unicron, a workload manager designed for efficient self-healing in large-scale language model training. Unicron optimizes the training process by minimizing failure-related costs across multiple concurrent tasks within a cluster. Its key features include in-band error detection for real-time error identification without extra overhead, a dynamic cost-aware plan generation mechanism for optimal reconfiguration, and an efficient transition strategy to reduce downtime during state changes. Deployed on a 128-GPU distributed cluster, Unicron demonstrates up to a 1.9x improvement in training efficiency over state-of-the-art methods, significantly reducing failure recovery costs and enhancing the reliability of large-scale language model training."} {"id": "arxiv:2401.00115v3", "text": "Treating the $X(4140)$ as a compact $J^{PC}=1^{++}$ $cs\\bar{c}\\bar{s}$ state and using its mass as a reference scale, we systematically estimate the masses of doubly heavy tetraquark states $QQ\\bar{q}\\bar{q}$ where $Q=c,b$ and $q=u,d,s$. Their decay properties are studied with a simple rearrangement scheme. Based on our results, the lowest $I(J^P)=0(1^+)$ $bb\\bar{n}\\bar{n}$ state is a stable tetraquark about 20 MeV below the $\\bar{B}^*\\bar{B}$ threshold. The mass and width of the low-mass $0(1^+)$ $cc\\bar{n}\\bar{n}$ ($n=u,d$) tetraquark are compatible with the $T_{cc}(3875)^+$ observed by the LHCb Collaboration. The location of the lowest $0(0^+)$ and $0(1^+)$ $bc\\bar{n}\\bar{n}$ states are found to be close to the $\\bar{B}D$ and $\\bar{B}^*D$ thresholds, respectively. We hope that the predicted ratios between partial widths of different channels may be helpful to identify compact tetraquark states from future measurements."} {"id": "arxiv:2401.00170v1", "text": "This work introduces the L3Cube-MahaSocialNER dataset, the first and largest social media dataset specifically designed for Named Entity Recognition (NER) in the Marathi language. The dataset comprises 18,000 manually labeled sentences covering eight entity classes, addressing challenges posed by social media data, including non-standard language and informal idioms. Deep learning models, including CNN, LSTM, BiLSTM, and Transformer models, are evaluated on the individual dataset with IOB and non-IOB notations. The results demonstrate the effectiveness of these models in accurately recognizing named entities in Marathi informal text. The L3Cube-MahaSocialNER dataset offers user-centric information extraction and supports real-time applications, providing a valuable resource for public opinion analysis, news, and marketing on social media platforms. We also show that the zero-shot results of the regular NER model are poor on the social NER test set thus highlighting the need for more social NER datasets. The datasets and models are publicly available at https://github.com/l3cube-pune/MarathiNLP"} {"id": "arxiv:2401.00174v2", "text": "We study nonequilibrium spin dynamics in differentially rotating systems, deriving an effective Hamiltonian for conduction electrons in the comoving frame. In contrast to conventional spin current generation mechanisms that require vorticity, our theory describes spins and spin currents arising from differentially rotating systems regardless of vorticity. We demonstrate the generation of spin currents in differentially rotating systems, such as liquid metals with Taylor-Couette flow. Our alternative mechanism will be important in the development of nanomechanical spin devices."} {"id": "arxiv:2401.00102v2", "text": "Machine-learning datasets are typically characterized by measuring their size and class balance. However, there exists a richer and potentially more useful set of measures, termed S-entropy (similarity-sensitive entropy), that incorporate elements' frequencies and between-element similarities. Although these have been available in the R and Julia programming languages for other applications, they have not been as readily available in Python, which is widely used for machine learning, and are not easily applied to machine-learning-sized datasets without special coding considerations. To address these issues, we developed $\\textit{sentropy}$, a Python package that calculates S-entropy and is tailored to large datasets. $\\textit{sentropy}$ can calculate any of the frequency-sensitive measures of Hill's D-number framework and their similarity-sensitive counterparts. $\\textit{sentropy}$ also outputs measures that compare datasets. We first briefly review S-entropy, illustrating how it incorporates elements' frequencies and elements' pairwise similarities. We then describe $\\textit{sentropy}$'s key features and usage. We end with several examples - immunomics, metagenomics, computational pathology, and medical imaging - illustrating $\\textit{sentropy}$'s applicability across a range of dataset types and fields."} {"id": "arxiv:2401.00139v3", "text": "This paper introduces a causal attribution model to enhance the interpretability of large language models (LLMs) and improve their causal reasoning abilities via precise fine-tuning. Despite LLMs' proficiency in diverse tasks, their reasoning processes often remain black box, and thus restrict targeted enhancement. We propose a novel causal attribution model that utilizes \"do-operators\" for constructing interventional scenarios, allowing us to quantify the contribution of different components in LLMs's causal reasoning process systematically. By assessing the proposed attribution scores through causal discovery tasks across various domains, we demonstrate that LLMs' effectiveness in causal discovery heavily relies on provided context and domain-specific knowledge but can also utilize numerical data with limited calculations in correlation, not causation. This motivates the proposed fine-tuned LLM for pairwise causal discovery, effectively and correctly leveraging both knowledge and numerical information."} {"id": "arxiv:2401.00216v2", "text": "With Wilson quarks, on-shell O($a$) improvement of the lattice QCD action is achieved by including the Sheikholeslami-Wohlert term and two further operators of mass dimension 5, which amount to a mass-dependent rescaling of the bare parameters. We here focus on the rescaled bare coupling, $\\tilde{g}_0^2 = g_0^2(1 + b_{\\rm g} am_{\\rm q})$, and the determination of $b_{\\rm g}(g_0^2)$, which is currently only known to 1-loop order of perturbation theory. We derive suitable improvement conditions in the chiral limit and in a finite space-time volume and evaluate these for different gluonic observables, both with and without the gradient flow. The choice of $β$-values and the line of constant physics are motivated by the ALPHA collaboration's decoupling strategy to determine $α_s(m_Z)$. However, the improvement conditions and some insight into systematic effects may prove useful in other contexts, too."} {"id": "arxiv:2401.00225v1", "text": "Dysarthria speech contains the pathological characteristics of vocal tract and vocal fold, but so far, they have not yet been included in traditional acoustic feature sets. Moreover, the nonlinearity and non-stationarity of speech have been ignored. In this paper, we propose a feature enhancement algorithm for dysarthria speech called WHFEMD. It combines empirical mode decomposition (EMD) and fast Walsh-Hadamard transform (FWHT) to enhance features. With the proposed algorithm, the fast Fourier transform of the dysarthria speech is first performed and then followed by EMD to get intrinsic mode functions (IMFs). After that, FWHT is used to output new coefficients and to extract statistical features based on IMFs, power spectral density, and enhanced gammatone frequency cepstral coefficients. To evaluate the proposed approach, we conducted experiments on two public pathological speech databases including UA Speech and TORGO. The results show that our algorithm performed better than traditional features in classification. We achieved improvements of 13.8% (UA Speech) and 3.84% (TORGO), respectively. Furthermore, the incorporation of an imbalanced classification algorithm to address data imbalance has resulted in a 12.18% increase in recognition accuracy. This algorithm effectively addresses the challenges of the imbalanced dataset and non-linearity in dysarthric speech and simultaneously provides a robust representation of the local pathological features of the vocal folds and tracts."} {"id": "arxiv:2401.00227v1", "text": "This paper investigates whether foreign investment (FDI) into Africa is at least partially responsive to World Bank-measured market friendliness. Specifically, I conducted analyses of four countries between 2009 and 2017, using cases that represent two of the highest scorers on the bank's Doing Business index as of 2008 (Mauritius and South Africa) and the two lowest scorers (DRC and CAR), and subsequently traced all four for growths or declines in FDI in relation to their scores in the index. The findings show that there is a moderate association between decreased costs of starting a business and growth of FDI. Mauritius, South Africa and the DRC reduced their total cost of starting a business by 71.7%, 143.7% and 122.9% for the entire period, and saw inward FDI increases of 167.6%, 79.8% and 152.21%, respectively. The CAR increased the cost of starting businesses but still saw increases in FDI. However, the country also saw the least amount of growth in FDI at only 13.3%."} {"id": "arxiv:2401.00229v2", "text": "Thermal leptogenesis is a mechanism that explains the observed asymmetry between matter and antimatter in the early universe. In this study, we review the impact of nonextensive Tsallis statistical mechanics on the early universe and study its effect on thermal leptogenesis. The study has found that the use of nonextensive statistical mechanics can affect the production of baryon asymmetry in thermal leptogenesis by modifying the equilibrium abundance of particles, decay, and washout parameters. Also, we show that nonextensive statistical mechanics potentially reduce the required right-handed neutrino mass scale."} {"id": "arxiv:2401.00230v2", "text": "In the domain of multivariate forecasting, transformer models stand out as powerful apparatus, displaying exceptional capabilities in handling messy datasets from real-world contexts. However, the inherent complexity of these datasets, characterized by numerous variables and lengthy temporal sequences, poses challenges, including increased noise and extended model runtime. This paper focuses on reducing redundant information to elevate forecasting accuracy while optimizing runtime efficiency. We propose a novel transformer forecasting framework enhanced by Principal Component Analysis (PCA) to tackle this challenge. The framework is evaluated by five state-of-the-art (SOTA) models and four diverse real-world datasets. Our experimental results demonstrate the framework's ability to minimize prediction errors across all models and datasets while significantly reducing runtime. From the model perspective, one of the PCA-enhanced models: PCA+Crossformer, reduces mean square errors (MSE) by 33.3% and decreases runtime by 49.2% on average. From the dataset perspective, the framework delivers 14.3% MSE and 76.6% runtime reduction on Electricity datasets, as well as 4.8% MSE and 86.9% runtime reduction on Traffic datasets. This study aims to advance various SOTA models and enhance transformer-based time series forecasting for intricate data. Code is available at: https://github.com/jingjing-unilu/PCA_Transformer."} {"id": "arxiv:2401.00255v1", "text": "The Wilcoxon signed-rank test and the Wilcoxon-Mann-Whitney test are commonly employed in one sample and two sample mean tests for one-dimensional hypothesis problems. For high-dimensional mean test problems, we calculate the asymptotic distribution of the maximum of rank statistics for each variable and suggest a max-type test. This max-type test is then merged with a sum-type test, based on their asymptotic independence offered by stationary and strong mixing assumptions. Our numerical studies reveal that this combined test demonstrates robustness and superiority over other methods, especially for heavy-tailed distributions."} {"id": "arxiv:2401.00292v1", "text": "The Multi-Objective Mixed-Integer Programming (MOMIP) problem is one of the most challenging. To derive its Pareto optimal solutions one can use the well-known Chebyshev scalarization and Mixed-Integer Programming (MIP) solvers. However, for a large-scale instance of the MOMIP problem, its scalarization may not be solved to optimality, even by state-of-the-art optimization packages, within the time limit imposed on the optimization. If a MIP solver cannot derive the optimal solution within the assumed time limit, it provides the optimality gap, which gauges the quality of the approximate solution. However, for the MOMIP case, no information is provided on the lower and upper bounds of the components of the Pareto optimal outcome. For the MOMIP problem with two and three objective functions, an algorithm is proposed to provide the so-called interval representation of the Pareto optimal outcome designated by the weighting vector when there is a time limit on solving the Chebyshev scalarization. Such interval representations can be used to navigate on the Pareto front. The results of several numerical experiments on selected large-scale instances of the multi-objective, multidimensional 0-1 knapsack problem illustrate the proposed approach. The limitations and possible enhancements of the proposed method are also discussed."} {"id": "arxiv:2401.00272v1", "text": "Proactively and naturally guiding the dialog from the non-recommendation context (e.g., Chit-chat) to the recommendation scenario (e.g., Music) is crucial for the Conversational Recommender System (CRS). Prior studies mainly focus on planning the next dialog goal~(e.g., chat on a movie star) conditioned on the previous dialog. However, we find the dialog goals can be simultaneously observed at different levels, which can be utilized to improve CRS. In this paper, we propose Dual-space Hierarchical Learning (DHL) to leverage multi-level goal sequences and their hierarchical relationships for conversational recommendation. Specifically, we exploit multi-level goal sequences from both the representation space and the optimization space. In the representation space, we propose the hierarchical representation learning where a cross attention module derives mutually enhanced multi-level goal representations. In the optimization space, we devise the hierarchical weight learning to reweight lower-level goal sequences, and introduce bi-level optimization for stable update. Additionally, we propose a soft labeling strategy to guide optimization gradually. Experiments on two real-world datasets verify the effectiveness of our approach. Code and data are available here."} {"id": "arxiv:2401.00265v1", "text": "In condensed matter physics, the Kagome lattice and its inherent flat bands have attracted considerable attention for their potential to host a variety of exotic physical phenomena. Despite extensive efforts to fabricate thin films of Kagome materials aimed at modulating the flat bands through electrostatic gating or strain manipulation, progress has been limited. Here, we report the observation of a novel $d$-orbital hybridized Kagome-derived flat band in Ag/Si(111) $\\sqrt{3}\\times\\sqrt{3}$ as revealed by angle-resolved photoemission spectroscopy. Our findings indicate that silver atoms on a silicon substrate form a Kagome-like structure, where a delicate balance in the hopping parameters of the in-plane $d$-orbitals leads to destructive interference, resulting in a flat band. These results not only introduce a new platform for Kagome physics but also illuminate the potential for integrating metal-semiconductor interfaces into Kagome-related research, thereby opening a new avenue for exploring ideal two-dimensional Kagome systems."} {"id": "arxiv:2401.00290v1", "text": "We consider the problem of red teaming LLMs on elementary calculations and algebraic tasks to evaluate how various prompting techniques affect the quality of outputs. We present a framework to procedurally generate numerical questions and puzzles, and compare the results with and without the application of several red teaming techniques. Our findings suggest that even though structured reasoning and providing worked-out examples slow down the deterioration of the quality of answers, the gpt-3.5-turbo and gpt-4 models are not well suited for elementary calculations and reasoning tasks, also when being red teamed."} {"id": "arxiv:2401.00220v3", "text": "Atacama Large Millimeter/Submillimeter Array (ALMA) has revolutionized the field of dust polarization in protoplanetary disks across multiple wavelengths. Previous observations and empirical modeling suggested multiple mechanisms of dust polarization toward HL Tau, including grain alignment and dust scattering. However, a detailed modeling of dust polarization based on grain alignment physics is not yet available. Here, using our updated POLARIS code, we perform numerical modeling of dust polarization arising from both grain alignment by Magnetically Enhanced Radiative Torque (MRAT) mechanism and self-scattering to reproduce the HL Tau polarization observed at three wavelengths 0.87, 1.3, and 3.1$\\,$mm. Our modeling results show that the observed multi-wavelength polarization could be reproduced only when large grains contain embedded iron inclusions and those with slow internal relaxation must have wrong internal alignment (i.e., the grain's major axis parallel to its angular momentum). The abundance of iron embedded inside grains in the form of clusters is constrained to be $\\gtrsim 16$%, and the number of iron atoms per cluster is $N_{\\rm cl} \\sim 9\\times10^2$. Maximum grain sizes probed at wavelengths $λ$ = 0.87, 1.3, and 3.1$\\,$mm are constrained at $\\sim$ 60, 80, and 90$\\,μ$m, respectively."} {"id": "arxiv:2401.00240v2", "text": "We revisit the study of the violation of the Leggett-Garg inequality in neutrino oscillation data as a mean to test some of the fundamental aspects of quantum mechanics. In particular, we consider the results of the Daya Bay and RENO reactor experiments, and the MINOS and NOvA accelerator experiments. We find that DB and MINOS exhibit a strong manifestation of Leggett-Garg violation, whereas for RENO and NOvA data, the indication is weaker. Considering the particular baselines and energy ranges explored by each experiment, our results demonstrate that the Leggett-Garg violation is more evident for smaller baseline-to-energy ratios in all the data sets studied, a relevant aspect to consider when looking for evidence of quantum mechanical decoherence in neutrino oscillations."} {"id": "arxiv:2401.00252v2", "text": "Motivated by the construction of Newton--Okounkov bodies and toric degenerations via cluster algebras in [GHKK18, FO25], we consider a family of Newton--Okounkov polytopes of a complex smooth Fano variety $X$ related by a composition of tropicalized cluster mutations. According to the work of [HK15], the toric degeneration associated with each Newton--Okounkov polytope $Δ$ in the family produces a completely integrable system of $X$ over $Δ$. We investigate circumstances in which each completely integrable system possesses a monotone Lagrangian torus fiber. We provide a sufficient condition, based on the data of tropical integer points and exchange matrices, for the family of constructed monotone Lagrangian tori to contain infinitely many monotone Lagrangian tori, no two of which are related by any symplectomorphism. By employing this criterion and exploiting the correspondence between the tropical integer points and the dual canonical basis elements, we generate infinitely many distinct monotone Lagrangian tori on flag manifolds of arbitrary type except in a few cases."} {"id": "arxiv:2401.00262v2", "text": "We study a generalized Witten's finiteness conjecture for the skein modules of oriented compact 3-manifolds with boundary. We formulate an equivalent version of the generalized finiteness conjecture using handlebodies and 2-handles, and prove the conjecture for some classes with the handlebodies of genus 2 and 3 using the equivalent version."} {"id": "arxiv:2401.00205v1", "text": "Online behavioral advertising (OBA) has a significant role in the digital economy. It allows advertisers to target consumers categorized according to their algorithmically inferred interests based on their behavioral data. As Alphabet and Meta gatekeep the Internet with their digital platforms and channel most of the consumer attention online, they are best placed to execute OBA and earn profits far exceeding fair estimations. There are increasing concerns that gatekeepers achieve such profitability at the expense of consumers, advertisers, and publishers who are dependent on their services to access the Internet. In particular, some claim that OBA systematically exploits consumers' decision-making vulnerabilities, creating internet infrastructure and relevant markets that optimize for consumer manipulation. Intuitively, consumer manipulation via OBA comes in tension with the ideal of consumer autonomy in liberal democracies. Nevertheless, academia has largely overlooked this phenomenon and instead has primarily focused on privacy and discrimination concerns of OBA. This article redirects academic discourse and regulatory focus on consumer manipulation via OBA. In doing so, first, this article elaborates on how OBA works. Second, it constructs an analytic framework for understanding manipulation. Third, it applies the theory of manipulation to OBA. As a result, this article illustrates the extent to which OBA leads to consumer manipulation. Crucially, this article is purely analytic and avoids normative evaluation of consumer manipulation via OBA. Evaluating consumer manipulation harms of OBA is an equally important but separate task and is pursued in another publication."} {"id": "arxiv:2401.00221v2", "text": "Patient-to-room assignment (PRA) is a scheduling problem in decision support for hospitals. It consists of assigning patients to rooms according to certain objectives, e.g., avoiding transfers and respecting single-room requests. This work presents combinatorial insights about the feasibility of PRA and about the assignment of patients to single rooms. We further compare different IP-formulations for PRA as well as the influence of different objectivs on the runtime. Based on these results, we develop a fast IP-based solution approach which obtains high quality solution. The applicability is verified through a computational study with instances derived from real-world data. Results indicate that large, real world instances can be solved to a high degree of optimality within (fractions of) seconds."} {"id": "arxiv:2401.00232v2", "text": "Low-dose emission tomography (ET) plays a crucial role in medical imaging, enabling the acquisition of functional information for various biological processes while minimizing the patient dose. However, the inherent randomness in the photon counting process is a source of noise which is amplified in low-dose ET. This review article provides an overview of existing post-processing techniques, with an emphasis on deep neural network (NN) approaches. Furthermore, we explore future directions in the field of NN-based low-dose ET. This comprehensive examination sheds light on the potential of deep learning in enhancing the quality and resolution of low-dose ET images, ultimately advancing the field of medical imaging."} {"id": "arxiv:2401.00237v1", "text": "Wind turbines are subjected to continuous rotational stresses and unusual external forces such as storms, lightning, strikes by flying objects, etc., which may cause defects in turbine blades. Hence, it requires a periodical inspection to ensure proper functionality and avoid catastrophic failure. The task of inspection is challenging due to the remote location and inconvenient reachability by human inspection. Researchers used images with cropped defects from the wind turbine in the literature. They neglected possible background biases, which may hinder real-time and autonomous defect detection using aerial vehicles such as drones or others. To overcome such challenges, in this paper, we experiment with defect detection accuracy by having the defects with the background using a two-step deep-learning methodology. In the first step, we develop virtual models of wind turbines to synthesize the near-reality images for four types of common defects - cracks, leading edge erosion, bending, and light striking damage. The Unity perception package is used to generate wind turbine blade defects images with variations in background, randomness, camera angle, and light effects. In the second step, a customized U-Net architecture is trained to classify and segment the defect in turbine blades. The outcomes of U-Net architecture have been thoroughly tested and compared with 5-fold validation datasets. The proposed methodology provides reasonable defect detection accuracy, making it suitable for autonomous and remote inspection through aerial vehicles."} {"id": "arxiv:2401.00253v2", "text": "Let $[n]:=\\lbrace 1,2,\\ldots,n \\rbrace$, and $M$ be a set of positive integers. Denote the family of all subsets of $[n]$ with sizes in $M$ by $\\binom{\\left[n\\right]}{M}$. The non-empty families $\\mathcal{A}\\subseteq\\binom{\\left[n\\right]}{R}$ and $\\mathcal{B}\\subseteq \\binom{\\left[n\\right]}{S}$ are said to be cross $t$-intersecting if $|A\\cap B|\\geq t$ for all $A\\in \\mathcal{A}$ and $B\\in \\mathcal{B}$. In this paper, we determine the maximum sum of sizes of non-empty cross $t$-intersecting families, and characterize the extremal families. Similar result for finite vector spaces is also proved."} {"id": "arxiv:2401.00257v1", "text": "There is a growing interest in the analysis of replication studies of original findings across many disciplines. When testing a hypothesis for an effect size, two Bayesian approaches stand out for their principled use of the Bayes factor (BF), namely the replication BF and the skeptical BF. In particular, the latter BF is based on the skeptical prior, which represents the opinion of an investigator who is unconvinced by the original findings and wants to challenge them. We embrace the skeptical perspective, and elaborate a novel mixture prior which incorporates skepticism while at the same time controlling for prior-data conflict within the original data. Consistency properties of the resulting skeptical mixture BF are provided together with an extensive analysis of the main features of our proposal. Finally, we apply our methodology to data from the Social Sciences Replication Project. In particular we show that, for some case studies where prior-data conflict is an issue, our method uses a more realistic prior and leads to evidence-classification for replication success which differs from the standard skeptical approach."} {"id": "arxiv:2401.00278v2", "text": "The intrinsic alignment (IA) of galaxies acts as a systematic effect in weak lensing measurements and tends to introduce biases. It mimics the gravitational lensing signal which makes it difficult to distinguish it from the true gravitational weak lensing effect. Hence, it is critical to account for the noise for correctly interpreting the results. This study aims at a quantitative analysis of IA using the Tidal Alignment and Tidal Torquing (TATT) model. We also investigate how the signals for shear and galaxy-galaxy lensing behave upon changing the parameters of the TATT model. The data for this study was prepared with a computational pipeline based on the Cocoa model to explore the parameter space of the intrinsic shape signal. Through this work, we identify that linear terms of the intrinsic shape signal are dominant in the case of GGL while the higher-order terms dictate the shear signal."} {"id": "arxiv:2401.00283v2", "text": "This article presents a comprehensive study on the emerging near-space communications (NS-COM) within the context of space-air-ground-sea integrated network (SAGSIN). Specifically, we firstly explore the recent technical developments of NS-COM, followed by the discussions about motivations behind integrating NS-COM into SAGSIN. To further demonstrate the necessity of NS-COM, a comparative analysis between the NS-COM network and other counterparts in SAGSIN is conducted, covering aspects of deployment, coverage, channel characteristics and unique problems of NS-COM network. Afterwards, the technical aspects of NS-COM, including channel modeling, random access, channel estimation, array-based beam management and joint network optimization, are examined in detail. Furthermore, we explore the potential applications of NS-COM, such as structural expansion in SAGSIN communication, civil aviation communication, remote and urgent communication, weather monitoring and carbon neutrality. Finally, some promising research avenues are identified, including stratospheric satellite (StratoSat) -to-ground direct links for mobile terminals, reconfigurable multiple-input multiple-output (MIMO) and holographic MIMO, federated learning in NS-COM networks, maritime communication, electromagnetic spectrum sensing and adversarial game, integrated sensing and communications, StratoSat-based radar detection and imaging, NS-COM assisted enhanced global navigation system, NS-COM assisted intelligent unmanned system and free space optical (FSO) communication. Overall, this paper highlights that the NS-COM plays an indispensable role in the SAGSIN puzzle, providing substantial performance and coverage enhancement to the traditional SAGSIN architecture."} {"id": "arxiv:2401.00286v1", "text": "The evolution of cybersecurity has spurred the emergence of autonomous threat hunting as a pivotal paradigm in the realm of AI-driven threat intelligence. This review navigates through the intricate landscape of autonomous threat hunting, exploring its significance and pivotal role in fortifying cyber defense mechanisms. Delving into the amalgamation of artificial intelligence (AI) and traditional threat intelligence methodologies, this paper delineates the necessity and evolution of autonomous approaches in combating contemporary cyber threats. Through a comprehensive exploration of foundational AI-driven threat intelligence, the review accentuates the transformative influence of AI and machine learning on conventional threat intelligence practices. It elucidates the conceptual framework underpinning autonomous threat hunting, spotlighting its components, and the seamless integration of AI algorithms within threat hunting processes.. Insightful discussions on challenges encompassing scalability, interpretability, and ethical considerations in AI-driven models enrich the discourse. Moreover, through illuminating case studies and evaluations, this paper showcases real-world implementations, underscoring success stories and lessons learned by organizations adopting AI-driven threat intelligence. In conclusion, this review consolidates key insights, emphasizing the substantial implications of autonomous threat hunting for the future of cybersecurity. It underscores the significance of continual research and collaborative efforts in harnessing the potential of AI-driven approaches to fortify cyber defenses against evolving threats."} {"id": "arxiv:2401.00271v1", "text": "Existing gait recognition benchmarks mostly include minor clothing variations in the laboratory environments, but lack persistent changes in appearance over time and space. In this paper, we propose the first in-the-wild benchmark CCGait for cloth-changing gait recognition, which incorporates diverse clothing changes, indoor and outdoor scenes, and multi-modal statistics over 92 days. To further address the coupling effect of clothing and viewpoint variations, we propose a hybrid approach HybridGait that exploits both temporal dynamics and the projected 2D information of 3D human meshes. Specifically, we introduce a Canonical Alignment Spatial-Temporal Transformer (CA-STT) module to encode human joint position-aware features, and fully exploit 3D dense priors via a Silhouette-guided Deformation with 3D-2D Appearance Projection (SilD) strategy. Our contributions are twofold: we provide a challenging benchmark CCGait that captures realistic appearance changes across an expanded and space, and we propose a hybrid framework HybridGait that outperforms prior works on CCGait and Gait3D benchmarks. Our project page is available at https://github.com/HCVLab/HybridGait."} {"id": "arxiv:2401.00296v2", "text": "When applying T-duality to a generic, non-extreme Killing horizon, T-duality is spacelike on one side and timelike on the other. We show, using simple examples from four-dimensional Einstein-Maxwell theory, that the image of the horizon is a singularity which can be understood as an interface between two different T-dual theories and their solutions. Using an embedding into type-II string theory, we show that the singularity occurs when scalars reach the boundary of moduli space, resulting in a breakdown of the effective field theory due to the presence of tensionless strings."} {"id": "arxiv:2401.00245v2", "text": "Among the variety of statistical intervals, highest-density regions (HDRs) stand out for their ability to effectively summarize a distribution or sample, unveiling its distinctive and salient features. An HDR represents the minimum size set that satisfies a certain probability coverage, and current methods for their computation require knowledge or estimation of the underlying probability distribution or density $f$. In this work, we illustrate a broader framework for computing HDRs, which generalizes the classical density quantile method introduced in the seminal paper of Hyndman (1996). The framework is based on neighbourhood measures, i.e., measures that preserve the order induced in the sample by $f$, and include the density $f$ as a special case. We explore a number of suitable distance-based measures, such as the $k$-nearest neighborhood distance, and some probabilistic variants based on copula models. An extensive comparison is provided, showing the advantages of the copula-based strategy, especially in those scenarios that exhibit complex structures (e.g., multimodalities or particular dependencies). Finally, we discuss the practical implications of our findings for estimating HDRs in real-world applications."} {"id": "arxiv:2401.00233v2", "text": "A model of a conducting cylinder with a radial temperature gradient which creates an electric field that increases with time in the surrounding vacuum is examined. The conditions under which this model functions are pointed out. An electric field is also generated when a magnetic field exists along the axis of the cylinder. This article discusses the interactions of the thermal flux, magnetic field, and charge distribution. Four models are considered with different conditions for the supply of electrons from a central source and the possibility either of capturing electrons inside the cylinder or their freely leaving it through the outer boundary."} {"id": "arxiv:2401.00250v3", "text": "For the first time, we estimate the in-medium mass shift of the two-flavored heavy mesons $B_c, B_c^*, B_s, B_s^*, D_s$ and $D_s^*$ in symmetric nuclear matter. The estimates are made by evaluating the lowest order one-loop self-energies. The enhanced excitations of intermediate state heavy-light mesons in symmetric nuclear matter are the origin of their negative mass shift. This negative mass shift may be regarded as a signature of partial restoration of chiral symmetry in an empirical sense because the origin of the negative mass shift in the study is not directly related to the chiral symmetry mechanism. Our results show that the magnitude of the mass shift for the $B_c$ meson ($\\bar{b} c$ or $b \\bar{c}$) is larger than those of the $η_c (\\bar{c} c)$ and $η_b (\\bar{b} b)$, different from a naive expectation that it would be in between them. While, that of the $B_c^*$ shows the in between of the $J/ψ$ and $Υ$. We observe that the lighter vector meson excitation in each meson self-energy gives a dominant contribution for the corresponding meson mass shift, $B_c, B_s,$ and $D_s$."} {"id": "arxiv:2401.00204v6", "text": "We present the electromagnetic (EM) dipole radiation flux from an eccentric Keplerian binary endowed with scalar charges, in the presence of scalar-photon coupling $φA_μA^μ$ or $φF_{μν}F^{μν}$. The scalar radiation is suppressed for orbital frequency below the scalar mass, while the scalar-mediated indirect EM radiation survives. We examine the constraints imposed on the scalar-photon and scalar-charge couplings by the current observational data of pulsar binaries, in case that the scalar charge is given by the muon number. The general extensions of the calculation to the quadrupole order and hyperbolic orbit are also discussed."} {"id": "arxiv:2401.00274v1", "text": "Recently, fictitious identical particles have provided a promising way to overcome the fermion sign problem and have been used in path integral Monte Carlo (PIMC) to accurately simulate warm dense matter with up to 1000 electrons (T. Dornheim et al., arXiv:2311.08098 (2023)). The inclusion of fictitious identical particles in path integral molecular dynamics (PIMD) can provide another way to simulate fermion systems. In a recent paper (J. Chem. Phys. 159, 154107 (2023)), Feldman and Hirshberg improved the recursive formula for PIMD of N identical bosons, significantly reducing the computational complexity from $O(PN^3)$ to $O(N^2+PN)$. In this paper, we extend this latest recursive formula for bosons to PIMD of fictitious identical particles to improve the efficiency of simulating fermion systems. We also provide the virial estimator for calculating energy by using the recursive technique. As an example, we use the quadratic scaling PIMD for fictitious identical particles to study the simulation of hundreds of fermions in a two-dimensional periodic potential, in the hope of providing a simulation tool for two-dimensional Fermi-Hubbard model and other strongly correlated fermion systems, such as the simulation of ultracold fermionic gases in optical lattices."} {"id": "arxiv:2401.00247v1", "text": "Numerical simulations can model the physical processes that govern cardiovascular device deployment. When such simulations incorporate digital twins; computational models of patient-specific anatomy, they can expedite and de-risk the device design process. Nonetheless, the exclusive use of patient-specific data constrains the anatomic variability which can be precisely or fully explored. In this study, we investigate the capacity of Latent Diffusion Models (LDMs) to edit digital twins to create anatomic variants, which we term digital siblings. Digital twins and their corresponding siblings can serve as the basis for comparative simulations, enabling the study of how subtle anatomic variations impact the simulated deployment of cardiovascular devices, as well as the augmentation of virtual cohorts for device assessment. However, while diffusion models have been characterized in their ability to edit natural images, their capacity to anatomically edit digital twins has yet to be studied. Using a case example centered on 3D digital twins of cardiac anatomy, we implement various methods for generating digital siblings and characterize them through morphological and topological analyses. We specifically edit digital twins to introduce anatomic variation at different spatial scales and within localized regions, demonstrating the existence of bias towards common anatomic features. We further show that such anatomic bias can be leveraged for virtual cohort augmentation through selective editing, partially alleviating issues related to dataset imbalance and lack of diversity. Our experimental framework thus delineates the limits and capabilities of using latent diffusion models in synthesizing anatomic variation for in silico trials."} {"id": "arxiv:2401.00215v4", "text": "Let $p$ and $q$ be two distinct odd primes, $p