arXiv ScienceSearch

arXiv subjects

Pedro Silva

Publications and source records attributed to Pedro Silva.

At least 19 recordsLinked to original sources

The $R$-Process Alliance: The $R$-Process Enhancement of Stars from Chemodynamically Tagged Groups in the Milky Way Halo

As part of the ongoing work of the $R$-Process Alliance (RPA), detailed abundance measurements of 29 heavy elements in three metal-poor stars, 2MASS J14592981$-$3852558, 2MASS J19445483$-$4039459, and 2MASS J15211026$-$0607566, are presented based on an analysis of high-resolution ($R\sim 80,000$), high signal-to-noise ``portrait'' spectra from the Magellan Inamori Kyocera Echelle (MIKE) spectrograph on the Magellan-Clay Telescope at Las Campanas Observatory. The selected targets were identified as $r$-process-enhanced metal-poor stars in previous RPA snapshot analyses. They have also been linked to possible chemodynamically tagged groups, indicating that the stars may have formed in dwarf galaxies that were later accreted into the Milky Way halo. These stars have also been tentatively linked to the Thamnos structure. The detailed chemical abundances in this work confirm that 2MASS J14592981$-$3852558 and J15211026$-$0607566 are $r$-II stars, while 2MASS J19445483$-$4039459 is found to lie just below the threshold for $r$-I status. The $r$-II stars show signs of slight enhancement in fission fragments compared to 2MASS J19445483$-$4039459. Based on radioactive age dating with Th, the $r$-process material in the two $r$-II stars is found to be old (with ages $>10$ Gyr); neither star shows signs of an actinide boost. The varying elemental compositions suggest that these stars likely did not originate in the same environment, though each could be consistent with originating in the Thamnos progenitor.

astro-ph.GA

Benford's Law as a Distributional Prior for Post-Training Quantization of Large Language Models

The rapid growth of Large Language Models (LLMs) intensifies the need for effective compression, with weight quantization being the most widely adopted technique. Standard uniform quantizers assume that parameters are evenly distributed, an assumption at odds with the highly skewed distributions observed in practice. We propose Benford-Quant, a simple, data-free non-uniform quantizer inspired by Benford's Law, which predicts that leading digits follow a logarithmic distribution. Benford-Quant replaces the uniform grid with a log-spaced codebook, dedicating more resolution to the frequent small-magnitude weights. We provide both theoretical intuition and empirical evidence: (i) weights in transformer transformational layers adhere closely to Benford statistics, while normalization layers systematically deviate; (ii) on Small Language Models (SLMs), Benford-Quant consistently improves perplexity, reducing 4-bit perplexity on Gemma-270M by more than 10%; and (iii) on larger LLMs, it remains competitive, with differences explained by over-parameterization effects. Our results indicate that incorporating a Benford-inspired prior into quantization grids is a low-cost modification that yields accuracy gains in aggressive few-bit regimes. Although it is not able to surpass the state of the art in tasks such as perplexity and LAMBADA, the Benford-Quant approach can be hybridized with other quantization methods-such as SmoothQuant and Activation-Aware Quantization-without major pipeline modification, potentially improving their performance.

cs.LG

PD-Loss: Proxy-Decidability for Efficient Metric Learning

Deep Metric Learning (DML) aims to learn embedding functions that map semantically similar inputs to proximate points in a metric space while separating dissimilar ones. Existing methods, such as pairwise losses, are hindered by complex sampling requirements and slow convergence. In contrast, proxy-based losses, despite their improved scalability, often fail to optimize global distribution properties. The Decidability-based Loss (D-Loss) addresses this by targeting the decidability index (d') to enhance distribution separability, but its reliance on large mini-batches imposes significant computational constraints. We introduce Proxy-Decidability Loss (PD-Loss), a novel objective that integrates learnable proxies with the statistical framework of d' to optimize embedding spaces efficiently. By estimating genuine and impostor distributions through proxies, PD-Loss combines the computational efficiency of proxy-based methods with the principled separability of D-Loss, offering a scalable approach to distribution-aware DML. Experiments across various tasks, including fine-grained classification and face verification, demonstrate that PD-Loss achieves performance comparable to that of state-of-the-art methods while introducing a new perspective on embedding optimization, with potential for broader applications.

cs.CV

Deep Learning for School Dropout Detection: A Comparison of Tabular and Graph-Based Models for Predicting At-Risk Students

Student dropout is a significant challenge in educational systems worldwide, leading to substantial social and economic costs. Predicting students at risk of dropout allows for timely interventions. While traditional Machine Learning (ML) models operating on tabular data have shown promise, Graph Neural Networks (GNNs) offer a potential advantage by capturing complex relationships inherent in student data if structured as graphs. This paper investigates whether transforming tabular student data into graph structures, primarily using clustering techniques, enhances dropout prediction accuracy. We compare the performance of GNNs (a custom Graph Convolutional Network (GCN) and GraphSAGE) on these generated graphs against established tabular models (Random Forest (RF), XGBoost, and TabNet) using a real-world student dataset. Our experiments explore various graph construction strategies based on different clustering algorithms (K-Means, HDBSCAN) and dimensionality reduction techniques (Principal Component Analysis (PCA), Uniform Manifold Approximation and Projection (UMAP)). Our findings demonstrate that a specific GNN configuration, GraphSAGE on a graph derived from PCA-KMeans clustering, achieved superior performance, notably improving the macro F1-score by approximately 7 percentage points and accuracy by nearly 2 percentage points over the strongest tabular baseline (XGBoost). However, other GNN configurations and graph construction methods did not consistently surpass tabular models, emphasizing the critical role of the graph generation strategy and GNN architecture selection. This highlights both the potential of GNNs and the challenges in optimally transforming tabular data for graph-based learning in this domain.

cs.LG

Top Quark at the New Physics Frontier

This Special Issue on "Top Quark at the New Physics Frontier" is devoted to the most massive fundamental elementary particle known, the top quark. The aim is to provide a comprehensive review of the current status and prospects of top quark physics at the Large Hadron Collider (LHC) and future colliders. We included articles that emphasize where the present understanding is incomplete and suggest new directions for research in this area.

hep-ph

A Systematic Review of ECG Arrhythmia Classification: Adherence to Standards, Fair Evaluation, and Embedded Feasibility

The classification of electrocardiogram (ECG) signals is crucial for early detection of arrhythmias and other cardiac conditions. However, despite advances in machine learning, many studies fail to follow standardization protocols, leading to inconsistencies in performance evaluation and real-world applicability. Additionally, hardware constraints essential for practical deployment, such as in pacemakers, Holter monitors, and wearable ECG patches, are often overlooked. Since real-world impact depends on feasibility in resource-constrained devices, ensuring efficient deployment is critical for continuous monitoring. This review systematically analyzes ECG classification studies published between 2017 and 2024, focusing on those adhering to the E3C (Embedded, Clinical, and Comparative Criteria), which include inter-patient paradigm implementation, compliance with Association for the Advancement of Medical Instrumentation (AAMI) recommendations, and model feasibility for embedded systems. While many studies report high accuracy, few properly consider patient-independent partitioning and hardware limitations. We identify state-of-the-art methods meeting E3C criteria and conduct a comparative analysis of accuracy, inference time, energy consumption, and memory usage. Finally, we propose standardized reporting practices to ensure fair comparisons and practical applicability of ECG classification models. By addressing these gaps, this study aims to guide future research toward more robust and clinically viable ECG classification systems.

cs.CV

Abundances of Neutron-Capture Elements in 62 Stars in the Globular Cluster Messier 15

M15 is a globular cluster with a known spread in neutron-capture elements. This paper presents abundances of neutron-capture elements for 62 stars in M15. Spectra were obtained with the Michigan/Magellan Fiber System (M2FS) spectrograph, covering a wavelength range from ~4430-4630 A. Spectral lines from Fe I, Fe II, Sr I, Zr II, Ba II, La II, Ce II, Nd II, Sm II, Eu II, and Dy II, were measured, enabling classifications and neutron-capture abundance patterns for the stars. Of the 62 targets, 44 are found to be highly Eu-enhanced r-II stars, another 17 are moderately Eu-enhanced r-I stars, and one star is found to have an s-process signature. The neutron-capture patterns indicate that the majority of the stars are consistent with enrichment by the r-process. The 62 target stars are found to show significant star-to-star spreads in Sr, Zr, Ba, La, Ce, Nd, Sm, Eu, and Dy, but no significant spread in Fe. The neutron-capture abundances are further found to have slight correlations with sodium abundances from the literature, unlike what has been previously found; follow-up studies are needed to verify this result. The findings in this paper suggest that the Eu-enhanced stars in M15 were enhanced by the same process, that the nucleosynthetic source of this Eu pollution was the r-process, and that the r-process source occurred as the first generation of cluster stars was forming.

astro-ph.SR

Progress in End-to-End Optimization of Detectors for Fundamental Physics with Differentiable Programming

In this article we examine recent developments in the research area concerning the creation of end-to-end models for the complete optimization of measuring instruments. The models we consider rely on differentiable programming methods and on the specification of a software pipeline including all factors impacting performance -- from the data-generating processes to their reconstruction and the extraction of inference on the parameters of interest of a measuring instrument -- along with the careful specification of a utility function well aligned with the end goals of the experiment. Building on previous studies originated within the MODE Collaboration, we focus specifically on applications involving instruments for particle physics experimentation, as well as industrial and medical applications that share the detection of radiation as their data-generating mechanism.

physics.ins-det

Representation Online Matters: Practical End-to-End Diversification in Search and Recommender Systems

As the use of online platforms continues to grow across all demographics, users often express a desire to feel represented in the content. To improve representation in search results and recommendations, we introduce end-to-end diversification, ensuring that diverse content flows throughout the various stages of these systems, from retrieval to ranking. We develop, experiment, and deploy scalable diversification mechanisms in multiple production surfaces on the Pinterest platform, including Search, Related Products, and New User Homefeed, to improve the representation of different skin tones in beauty and fashion content. Diversification in production systems includes three components: identifying requests that will trigger diversification, ensuring diverse content is retrieved from the large content corpus during the retrieval stage, and finally, balancing the diversity-utility trade-off in a self-adjusting manner in the ranking stage. Our approaches, which evolved from using Strong-OR logical operator to bucketized retrieval at the retrieval stage and from greedy re-rankers to multi-objective optimization using determinantal point processes for the ranking stage, balances diversity and utility while enabling fast iterations and scalable expansion to diversification over multiple dimensions. Our experiments indicate that these approaches significantly improve diversity metrics, with a neutral to a positive impact on utility metrics and improved user satisfaction, both qualitatively and quantitatively, in production. An accessible PDF of this article is available at https://drive.google.com/file/d/1p5PkqC-sdtX19Y_IAjZCtiSxSEX1IP3q/view

cs.IR

CapsProm: A Capsule Network For Promoter Prediction

Locating the promoter region in DNA sequences is of paramount importance in the field of bioinformatics. This is a problem widely studied in the literature, however, not yet fully resolved. Some researchers have presented remarkable results using convolution networks, that allowed the automatic extraction of features from a DNA chain. However, a universal architecture that could generalize to several organisms has not yet been achieved, and thus, requiring researchers to seek new architectures and hyperparameters for each new organism evaluated. In this work, we propose a versatile architecture, based on capsule network, that can accurately identify promoter sequences in raw DNA data from seven different organisms, eukaryotic, and prokaryotic. Our model, the CapsProm, could assist in the transfer of learning between organisms and expand its applicability. Furthermore the CapsProm showed competitive results, overcoming the baseline method in five out of seven of the tested datasets (F1-score). The models and source code are made available at https://github.com/lauromoraes/CapsNet-promoter.

cs.LG

A Decidability-Based Loss Function

Nowadays, deep learning is the standard approach for a wide range of problems, including biometrics, such as face recognition and speech recognition, etc. Biometric problems often use deep learning models to extract features from images, also known as embeddings. Moreover, the loss function used during training strongly influences the quality of the generated embeddings. In this work, a loss function based on the decidability index is proposed to improve the quality of embeddings for the verification routine. Our proposal, the D-loss, avoids some Triplet-based loss disadvantages such as the use of hard samples and tricky parameter tuning, which can lead to slow convergence. The proposed approach is compared against the Softmax (cross-entropy), Triplets Soft-Hard, and the Multi Similarity losses in four different benchmarks: MNIST, Fashion-MNIST, CIFAR10 and CASIA-IrisV4. The achieved results show the efficacy of the proposal when compared to other popular metrics in the literature. The D-loss computation, besides being simple, non-parametric and easy to implement, favors both the inter-class and intra-class scenarios.

cs.CV

Performing Creativity With Computational Tools

The introduction of new tools in people's workflow has always been promotive of new creative paths. This paper discusses the impact of using computational tools in the performance of creative tasks, especially focusing on graphic design. The study was driven by a grounded theory methodology, applied to a set of semi-structured interviews, made to twelve people working in the areas of graphic design, data science, computer art, music and data visualisation. Among other questions, the results suggest some scenarios in which it is or it is not worth investing in the development of new intelligent creativity-aiding tools.

cs.CY

Multirotors from Takeoff to Real-Time Full Identification Using the Modified Relay Feedback Test and Deep Neural Networks

Low cost real-time identification of multirotor unmanned aerial vehicle (UAV) dynamics is an active area of research supported by the surge in demand and emerging application domains. Such real-time identification capabilities shorten development time and cost, making UAVs' technology more accessible, and enable a wide variety of advanced applications. In this paper, we present a novel comprehensive approach, called DNN-MRFT, for real-time identification and tuning of multirotor UAVs using the Modified Relay Feedback Test (MRFT) and Deep Neural Networks (DNN). The main contribution is the development of a generalized framework for the application of DNN-MRFT to higher-order systems. One of the notable advantages of DNN-MRFT is the exact estimation of identified process gain, which mitigates the inaccuracies introduced due to the use of the describing function method in approximating the response of Lure's systems. A secondary contribution is a generalized controller based on DNN-MRFT that takes-off a UAV with unknown dynamics and identifies the inner loops dynamics in-flight. Using the developed framework, DNN-MRFT is sequentially applied to the outer translational loops of the UAV utilizing in-flight results obtained for the inner attitude loops. DNN-MRFT takes on average 15 seconds to get the full knowledge of multirotor UAV dynamics and without any further tuning or calibration the UAV would be able to pass through a vertical window, and accurately follow trajectories achieving state-of-the-art performance. Such demonstrated accuracy, speed, and robustness of identification pushes the limits of state-of-the-art in real-time identification of UAVs.

eess.SY

ARCHI: pipeline for light curve extraction of CHEOPS background star

High precision time series photometry from space is being used for a number of scientific cases. In this context, the recently launched CHEOPS (ESA) mission promises to bring 20 ppm precision over an exposure time of 6 hours, when targeting nearby bright stars, having in mind the detailed characterization of exoplanetary systems through transit measurements. However, the official CHEOPS (ESA) mission pipeline only provides photometry for the main target (the central star in the field). In order to explore the potential of CHEOPS photometry for all stars in the field, in this paper we present archi, an additional open-source pipeline module{\dag}to analyse the background stars present in the image. As archi uses the official Data Reduction Pipeline data as input, it is not meant to be used as independent tool to process raw CHEOPS data but, instead, to be used as an add-on to the official pipeline. We test archi using CHEOPS simulated images, and show that photometry of background stars in CHEOPS images is only slightly degraded (by a factor of 2 to 3) with respect to the main target. This opens a potential for the use of CHEOPS to produce photometric time series of several close-by targets at once, as well as to use different stars in the image to calibrate systematic errors. We also show one clear scientific application where the study of the companion light curve can be important for the understanding of the contamination on the main target.

astro-ph.IM

A Token-Based MAC Solution for WiLD Point-To-Multipoint Links

The inefficiency of the fundamental access method of the IEEEE 802.11 standard is a well-known problem in scenarios where multiple long range and faulty links compete for the shared medium. The alternatives found in the literature are mostly focused on point-to-point and rarely on long range links. Currently, there is no solution optimized for maritime scenarios, where the point-to-multipoint links can reach several tens of kilometers, while suffering from degradation due to the harsh environment. This paper presents a novel MAC protocol that uses explicit signalling messages to control access to the medium by a central node. In this solution, this role is played by an Access Point that sends token messages addressed to each associated station, which is granted exclusive access to the medium and assigned a number of credits; after sending its own packets (if any), a station must release control of the token. Simulation results show that this mechanism has superior performance in point-to-multipoint, long range and faulty links, allowing a fairly and more efficient usage of the shared resources among all nodes.

cs.NI

On the Wilson Monoid of a Pairwise Balanced Design

We give a new perspective of the relationship between simple matroids of rank 3 and pairwise balanced designs, connecting Wilson's theorems and tools with the theory of truncated boolean representable simplicial complexes. We also introduce the concept of Wilson monoid W(X) of a pairwise balanced design X. We present some general algebraic properties and study in detail the cases of Steiner triple systems up to 19 points, as well as the case where a single block has more than 2 elements

math.CO

Truncated Boolean Representable Simplicial Complexes

We extend, in significant ways, the theory of truncated boolean representable simplicial complexes introduced in 2015. This theory, which includes all matroids, represents the largest class of finite simplicial complexes for which combinatorial geometry can be meaningfully applied

math.CO

Improving Malware Detection Accuracy by Extracting Icon Information

Detecting PE malware files is now commonly approached using statistical and machine learning models. While these models commonly use features extracted from the structure of PE files, we propose that icons from these files can also help better predict malware. We propose an innovative machine learning approach to extract information from icons. Our proposed approach consists of two steps: 1) extracting icon features using summary statics, histogram of gradients (HOG), and a convolutional autoencoder, 2) clustering icons based on the extracted icon features. Using publicly available data and by using machine learning experiments, we show our proposed icon clusters significantly boost the efficacy of malware prediction models. In particular, our experiments show an average accuracy increase of 10% when icon clusters are used in the prediction model.

cs.CR