arXiv ScienceSearch

arXiv subjects

Jonathan Scott

Publications and source records attributed to Jonathan Scott.

18 recordsLinked to original sources

Personalized Product Search Ranking: A Multi-Task Learning Approach with Tabular and Non-Tabular Data

In this paper, we present a novel model architecture for optimizing personalized product search ranking using a multi-task learning (MTL) framework. Our approach uniquely integrates tabular and non-tabular data, leveraging a pre-trained TinyBERT model for semantic embeddings and a novel sampling technique to capture diverse customer behaviors. We evaluate our model against several baselines, including XGBoost, TabNet, FT-Transformer, DCN-V2, and MMoE, focusing on their ability to handle mixed data types and optimize personalized ranking. Additionally, we propose a scalable relevance labeling mechanism based on click-through rates, click positions, and semantic similarity, offering an alternative to traditional human-annotated labels. Experimental results show that combining non-tabular data with advanced embedding techniques in multi-task learning paradigm significantly enhances model performance. Ablation studies further underscore the benefits of incorporating relevance labels, fine-tuning TinyBERT layers, and TinyBERT query-product embedding interactions. These results demonstrate the effectiveness of our approach in achieving improved personalized product search ranking.

cs.IR

Differentially Private Federated $k$-Means Clustering with Server-Side Data

Clustering is a cornerstone of data analysis that is particularly suited to identifying coherent subgroups or substructures in unlabeled data, as are generated continuously in large amounts these days. However, in many cases traditional clustering methods are not applicable, because data are increasingly being produced and stored in a distributed way, e.g. on edge devices, and privacy concerns prevent it from being transferred to a central server. To address this challenge, we present FedDP-KMeans, a new algorithm for $k$-means clustering that is fully-federated as well as differentially private. Our approach leverages (potentially small and out-of-distribution) server-side data to overcome the primary challenge of differentially private clustering methods: the need for a good initialization. Combining our initialization with a simple federated DP-Lloyds algorithm we obtain an algorithm that achieves excellent results on synthetic and real-world benchmark tasks. We also provide a theoretical analysis of our method that provides bounds on the convergence speed and cluster identification success.

cs.CR

Federated Learning with Unlabeled Clients: Personalization Can Happen in Low Dimensions

Personalized federated learning has emerged as a popular approach to training on devices holding statistically heterogeneous data, known as clients. However, most existing approaches require a client to have labeled data for training or finetuning in order to obtain their own personalized model. In this paper we address this by proposing FLowDUP, a novel method that is able to generate a personalized model using only a forward pass with unlabeled data. The generated model parameters reside in a low-dimensional subspace, enabling efficient communication and computation. FLowDUP's learning objective is theoretically motivated by our new transductive multi-task PAC-Bayesian generalization bound, that provides performance guarantees for unlabeled clients. The objective is structured in such a way that it allows both clients with labeled data and clients with only unlabeled data to contribute to the training process. To supplement our theoretical results we carry out a thorough experimental evaluation of FLowDUP, demonstrating strong empirical performance on a range of datasets with differing sorts of statistically heterogeneous clients. Through numerous ablation studies, we test the efficacy of the individual components of the method.

cs.LG

The Simplicial Loop Space of a Simplicial Complex

Given a simplicial complex $X$, we construct a simplicial complex $\Omega X$ that may be regarded as a combinatorial version of the based loop space of a topological space. Our construction explicitly describes the simplices of $\Omega X$ directly in terms of the simplices of $X$. Working at a purely combinatorial level, we show two main results that confirm the (combinatorial) algebraic topology of our $\Omega X$ behaves like that of the topological based loop space. Whereas our $\Omega X$ is generally a disconnected simplical complex, each component of $\Omega X$ has the same edge group, up to isomorphism. We show an isomorphism between the edge group of $\Omega X$ and the combinatorial second homotopy group of $X$ as it has been defined in separate work (arxiv:2503.23651). Finally, we enter the topological setting and, relying on prior work of Stone, show a homotopy equivalence between the spatial realization of our $\Omega X$ and the based loop space of the spatial realization of $X$.

math.AT

Improved Modelling of Federated Datasets using Mixtures-of-Dirichlet-Multinomials

In practice, training using federated learning can be orders of magnitude slower than standard centralized training. This severely limits the amount of experimentation and tuning that can be done, making it challenging to obtain good performance on a given task. Server-side proxy data can be used to run training simulations, for instance for hyperparameter tuning. This can greatly speed up the training pipeline by reducing the number of tuning runs to be performed overall on the true clients. However, it is challenging to ensure that these simulations accurately reflect the dynamics of the real federated training. In particular, the proxy data used for simulations often comes as a single centralized dataset without a partition into distinct clients, and partitioning this data in a naive way can lead to simulations that poorly reflect real federated training. In this paper we address the challenge of how to partition centralized data in a way that reflects the statistical heterogeneity of the true federated clients. We propose a fully federated, theoretically justified, algorithm that efficiently learns the distribution of the true clients and observe improved server-side simulations when using the inferred distribution to create simulated clients from the centralized data.

cs.LG

PeFLL: Personalized Federated Learning by Learning to Learn

We present PeFLL, a new personalized federated learning algorithm that improves over the state-of-the-art in three aspects: 1) it produces more accurate models, especially in the low-data regime, and not only for clients present during its training phase, but also for any that may emerge in the future; 2) it reduces the amount of on-client computation and client-server communication by providing future clients with ready-to-use personalized models that require no additional finetuning or optimization; 3) it comes with theoretical guarantees that establish generalization from the observed clients to future ones. At the core of PeFLL lies a learning-to-learn approach that jointly trains an embedding network and a hypernetwork. The embedding network is used to represent clients in a latent descriptor space in a way that reflects their similarity to each other. The hypernetwork takes as input such descriptors and outputs the parameters of fully personalized client models. In combination, both networks constitute a learning algorithm that achieves state-of-the-art performance in several personalized federated learning benchmarks.

cs.LG

Cross-client Label Propagation for Transductive and Semi-Supervised Federated Learning

We present Cross-Client Label Propagation(XCLP), a new method for transductive federated learning. XCLP estimates a data graph jointly from the data of multiple clients and computes labels for the unlabeled data by propagating label information across the graph. To avoid clients having to share their data with anyone, XCLP employs two cryptographically secure protocols: secure Hamming distance computation and secure summation. We demonstrate two distinct applications of XCLP within federated learning. In the first, we use it in a one-shot way to predict labels for unseen test points. In the second, we use it to repeatedly pseudo-label unlabeled training data in a federated semi-supervised setting. Experiments on both real federated and standard benchmark datasets show that in both applications XCLP achieves higher classification accuracy than alternative approaches.

cs.LG

Achieving Reliable and Repeatable Electrochemical Impedance Spectroscopy of Rechargeable Batteries at Extra-Low Frequencies

There is a need for techniques for efficient and accurate measurement of the impedance of rechargeable batteries at extra-low frequencies (ELFs, typically below 10uHz), as these reflect real usage and cycling patterns, and their importance in fractional battery circuit modeling is becoming increasingly apparent. Major impediments include the time required to perform such measurements, and `drift' in impedance values when measurements are taken from the same battery at different times. Moreover, commercial impedance analyzers are generally unable to measure at frequencies of the order of microhertz. We describe here our use of programmable two-quadrant power supplies to deliver multiple small signal measurement tones in the presence of large signal `working' currents, and our use of these data to generate impedance measurements with good precision and in reasonable time.

physics.ins-det

Charge capacity characteristics of a Lithium Nickel-Cobalt-Aluminium Oxide battery show fractional-derivative behavior

Batteries experience capacity offset where available charge depends on the rate at which this charge is drawn. In this work we analyze the capacity offset of a 4.8 A h lithium nickel-cobalt-aluminium oxide battery using an equivalent circuit model of a fractional capacitor in series with a resistor. In this case, the available charge, in theory, becomes infinite in the limit of infinitesimal rate. We show that the fractional properties of the capacitor can be extracted from the charge against rate plot. We then use a network of RC elements to represent the fractional capacitor in order to simulate the data with Matlab. We find that the fractional exponent alpha obtained in this way, 0.971, agrees with that obtained in a more traditional manner from an impedance versus frequency plot, although the fractional capacity does not. Such an approach demonstrates the importance of a fractional description for capacity offset even when an element is nearly a pure capacitor and is valuable for predictions of state-of-charge when low currents are drawn.

eess.SY

Exact weights, path metrics, and algebraic Wasserstein distances

We use weights on objects in an abelian category to define what we call a path metric. We introduce three special classes of weight: those compatible with short exact sequences; those induced by their path metric; and those which bound their path metric. We prove that these conditions are in fact equivalent, and call such weights exact. As a special case of a path metric, we obtain a distance for generalized persistence modules whose indexing category is a measure space. We use this distance to define Wasserstein distances, which coincide with the previously defined Wasserstein distances for one-parameter persistence modules. For one-parameter persistence modules, we also describe maps to and from an interval module, and we give a matrix reduction for monomorphisms and epimorphisms.

math.RA

Interleaving and Gromov-Hausdorff distance

One of the central notions to emerge from the study of persistent homology is that of interleaving distance. It has found recent applications in symplectic and contact geometry, sheaf theory, computational geometry, and phylogenetics. Here we present a general study of this topic. We define interleaving of functors with common codomain as solutions to an extension problem. In order to define interleaving distance in this setting we are led to categorical generalizations of Hausdorff distance, Gromov-Hausdorff distance, and the space of metric spaces. We obtain comparisons with previous notions of interleaving via the study of future equivalences. As an application we recover a definition of shift equivalences of discrete dynamical systems.

math.CT

Metrics for generalized persistence modules

We consider the question of defining interleaving metrics on generalized persistence modules over arbitrary preordered sets. Our constructions are functorial, which implies a form of stability for these metrics. We describe a large class of examples, inverse-image persistence modules, which occur whenever a topological space is mapped to a metric space. Several standard theories of persistence and their stability can be described in this framework. This includes the classical case of sublevelset persistent homology. We introduce a distinction between `soft' and `hard' stability theorems. While our treatment is direct and elementary, the approach can be explained abstractly in terms of monoidal functors.

math.AT

Twisting structures and strongly homotopy morphisms

In an application of the notion of twisting structures introduced by Hess and Lack, we define twisted composition products of symmetric sequences of chain complexes that are degreewise projective and finitely generated. Let Q be a cooperad and let BP be the bar construction on the operad P. To each morphism of cooperads g from Q to BP is associated a P-co-ring, K(g), which generalizes the two-sided Koszul and bar constructions. When the co-unit from K(g) to P is a quasi-isomorphism, we show that the Kleisli category for K(g) is isomorphic to the category of P-algebras and of their morphisms up to strong homotopy, and we give the classifying morphisms for both strict and homotopy P-algebras. Parametrized morphisms of (co)associative chain (co)algebras up to strong homotopy are also introduced and studied, and a general existence theorem is proved. In the appendix, we study the particular case of the two-sided Koszul resolution of the associative operad.

math.AT

CoHochschild homology of chain coalgebras

Generalizing work of Doi and of Idrissi, we define a coHochschild homology theory for chain coalgebras over any commutative ring and prove its naturality with respect to morphisms of chain coalgebras up to strong homotopy. As a consequence we obtain that if the comultiplication of a chain coalgebra $C$ is itself a morphism of chain coalgebras up to strong homotopy, then the coHochschild complex $\cohoch (C)$ admits a natural comultiplicative structure. In particular, if $K$ is a reduced simplicial set and $C_{*}K$ is its normalized chain complex, then $\cohoch (C_{*}K)$ is naturally a homotopy-coassociative chain coalgebra. We provide a simple, explicit formula for the comultiplication on $\cohoch (C_{*}K)$ when $K$ is a simplicial suspension. The coHochschild complex construction is topologically relevant. Given two simplicial maps $g,h:K\to L$, where $K$ and $L$ are reduced, the homology of the coHochschild complex of $C_{*}L$ with coefficients in $C_{*}K$ is isomorphic to the homology of the homotopy coincidence space of the geometric realizations of $g$ and $h$, and this isomorphism respects comultiplicative structure. In particular, there a isomorphism, respecting comultiplicative structure, from the homology of $\cohoch(C_{*}K)$ to $H_{*}\op L|K|$, the homology of the free loops on the geometric realization of $K$.

math.AT

A chain coalgebra model for the James map

Let EK be the simplicial suspension of a pointed simplicial set K. We construct a chain model of the James map, $α_{K} : CK \to ΩCEK$. We compute the cobar diagonal on $ΩCEK$, not assuming that $EK$ is 1-reduced, and show that $α_{K}$ is comultiplicative. As a result, the natural isomorphism of chain algebras $TCK \cong ΩCK$ preserves diagonals. In an appendix, we show that the Milgram map, $Ω(A \otimes B) \to ΩA \otimes ΩB$, where A and B are coaugmented coalgebras, forms part of a strong deformation retract of chain complexes. Therefore, it is a chain equivalence even when A and B are not 1-connected.

math.AT

Co-rings over operads characterize morphisms

Let M be a bicomplete, closed symmetric monoidal category. Let P be an operad in M, i.e., a monoid in the category of symmetric sequences of objects in M, with its composition monoidal structure. Let R be a P-co-ring, i.e., a comonoid in the category of P-bimodules. The co-ring R induces a natural ``fattening'' of the category of P-(co)algebras, expanding the morphism sets while leaving the objects fixed. Co-rings over operads are thus ``relative operads,'' parametrizing morphisms as operads parametrize (co)algebras. Let A denote the associative operad in the category of chain complexes. We define a ``diffracting'' functor that produces A-co-rings from symmetric sequences of chain coalgebras, leading to a multitude of ``fattened'' categories of (co)associative chain (co)algebras. In particular, we obtain a purely operadic description of the categories DASH and DCSH first defined by Gugenheim and Munkholm, via an A-co-ring that has the two-sided Koszul resolution of A as its underlying A-bimodule. The diffracting functor plays a crucial role in enabling us to prove existence of higher, ``up to homotopy'' structure of morphisms via acyclic models methods. It has already been successfully applied in this sense in a number of recent articles and preprints.

math.AT

A canonical enriched Adams-Hilton model for simplicial sets

For any 1-reduced simplicial set $K$ we define a canonical, coassociative coproduct on $\Om C(K)$, the cobar construction applied to the normalized, integral chains on $K$, such that any canonical quasi-isomorphism of chain algebras from $\Om C(K)$ to the normalized, integral chains on $GK$, the loop group of $K$, is a coalgebra map up to strong homotopy. Our proof relies on the operadic description of the category of chain coalgebras and of strongly homotopy coalgebra maps given in math.AT/0505559.

math.AT