arXiv ScienceSearch

arXiv subjects

Dylan Cutler

Publications and source records attributed to Dylan Cutler.

6 recordsLinked to original sources

Byte by Byte: Unmasking Browser Fingerprinting at the Function Level Using V8 Bytecode Transformers

Browser fingerprinting enables persistent cross-site user tracking via subtle techniques that often evade conventional defenses or cause website breakage when script-level blocking countermeasures are applied. Addressing these challenges requires detection methods offering both function-level precision to minimize breakage and inherent robustness against code obfuscation and URL manipulation. We introduce ByteDefender, the first system leveraging V8 engine bytecode to detect fingerprinting operations specifically at the JavaScript function level. A Transformer-based classifier, trained offline on bytecode sequences, accurately identifies functions exhibiting fingerprinting behavior. We develop and evaluate light-weight signatures derived from this model to enable low-overhead, on-device matching against function bytecode during compilation but prior to execution, which only adds a 4% (average) latency to the page load time. This mechanism facilitates targeted, real-time prevention of fingerprinting function execution, thereby preserving legitimate script functionality. Operating directly on bytecode ensures inherent resilience against common code obfuscation and URL-based evasion. Our evaluation on the top 100k websites demonstrates high detection accuracy at both function- and script-level, with substantial improvements over state-of-the-art AST-based methods, particularly in robustness against obfuscation. ByteDefender offers a practical framework for effective, precise, and robust fingerprinting mitigation.

cs.CR

StagFormer: Time Staggering Transformer Decoding for RunningLayers In Parallel

Decoding in a Transformer based language model is inherently sequential as a token's embedding needs to pass through all the layers in the network before the generation of the next token can begin. In this work, we propose a new architecture StagFormer (Staggered Transformer), which staggers execution along the sequence axis and thereby enables parallelizing the decoding process along the depth of the model. We achieve this by breaking the dependency of the token representation at time step $i$ in layer $l$ upon the representations of tokens until time step $i$ from layer $l-1$. Instead, we stagger the execution and only allow a dependency on token representations until time step $i-1$. The later sections of the Transformer still get access to the "rich" representations from the prior section but only from those token positions which are one time step behind. StagFormer allows for different sections of the model to be executed in parallel yielding a potential speedup in decoding while being quality neutral in our simulations. We also explore many natural extensions of this idea. We present how weight-sharing across the different sections being staggered can be more practical in settings with limited memory. We explore the efficacy of using a bounded window attention to pass information from one section to another which helps drive further latency gains for some applications. We also explore the scalability of the staggering idea over more than 2 sections of the Transformer. Finally, we show how one can approximate a recurrent model during inference using weight-sharing. This variant can lead to substantial gains in quality for short generations while being neutral in its latency impact.

cs.LG

Alternating Updates for Efficient Transformers

It has been well established that increasing scale in deep transformer networks leads to improved quality and performance. However, this increase in scale often comes with prohibitive increases in compute cost and inference latency. We introduce Alternating Updates (AltUp), a simple-to-implement method to increase a model's capacity without the computational burden. AltUp enables the widening of the learned representation, i.e., the token embedding, while only incurring a negligible increase in latency. AltUp achieves this by working on a subblock of the widened representation at each layer and using a predict-and-correct mechanism to update the inactivated blocks. We present extensions of AltUp, such as its applicability to the sequence dimension, and demonstrate how AltUp can be synergistically combined with existing approaches, such as Sparse Mixture-of-Experts models, to obtain efficient models with even higher capacity. Our experiments on benchmark transformer models and language tasks demonstrate the consistent effectiveness of AltUp on a diverse set of scenarios. Notably, on SuperGLUE and SQuAD benchmarks, AltUp enables up to $87\%$ speedup relative to the dense baselines at the same accuracy.

cs.LG

Computational Framework for Behind-The-Meter DER Techno-Economic Modeling and Optimization -- REopt Lite

The global energy system is undergoing a major transformation. Renewable energy generation is growing and is projected to accelerate further with the global emphasis on decarbonization. Furthermore, distributed generation is projected to play a significant role in the new energy system, and energy models are playing a key role in understanding how distributed generation can be integrated reliably and economically. The deployment of massive amounts of distributed generation requires understanding the interface of technology, economics, and policy in the energy modeling process. In this work, we present an end-to-end computational framework for distributed energy resource (DER) modeling, REopt Lite which addresses this need effectively. We describe the problem space, the building blocks of the model, the scaling capabilities of the design, the optimization formulation, and the accessibility of the model. We present a framework for accelerating the techno-economic analysis of behind-the-meter distributed energy resources to enable rapid planning and decision-making, thereby significantly boosting the rate the renewable energy deployment. Lastly, but equally importantly, this computation framework is open-sourced to facilitate transparency, flexibility, and wider collaboration opportunities within the worldwide energy modeling community.

cs.OH

A Unified Architecture for Data-Driven Metadata Tagging of Building Automation Systems

This article presents a Unified Architecture for automated point tagging of Building Automation System data, based on a combination of data-driven approaches. Advanced energy analytics applications-including fault detection and diagnostics and supervisory control-have emerged as a significant opportunity for improving the performance of our built environment. Effective application of these analytics depends on harnessing structured data from the various building control and monitoring systems, but typical Building Automation System implementations do not employ any standardized metadata schema. While standards such as Project Haystack and Brick Schema have been developed to address this issue, the process of structuring the data, i.e., tagging the points to apply a standard metadata schema, has, to date, been a manual process. This process is typically costly, labor-intensive, and error-prone. In this work we address this gap by proposing a UA that automates the process of point tagging by leveraging the data accessible through connection to the BAS, including time series data and the raw point names. The UA intertwines supervised classification and unsupervised clustering techniques from machine learning and leverages both their deterministic and probabilistic outputs to inform the point tagging process. Furthermore, we extend the UA to embed additional input and output data-processing modules that are designed to address the challenges associated with the real-time deployment of this automation solution. We test the UA on two datasets for real-life buildings: 1. commercial retail buildings and 2. office buildings from the National Renewable Energy Laboratory campus. The proposed methodology correctly applied 85-90 percent and 70-75 percent of the tags in each of these test scenarios, respectively.

cs.CY

Evaluation of Centralized and Distributed Microgrid Topologies Considering Power Quality Constraints

Integration of renewable generation and energy storage technologies with conventional generation supports increased resilience, lower-costs, and clean energy goals. Traditionally, energy supply needs of rural off-grid communities have been addressed with diesel-generation. But with rapidly falling renewable generation costs, mini-grids are transforming into hybrid systems with a mix of renewables, energy storage, and diesel-generation. Optimal design of hybrid mini-grid requires an understanding of both the economic and power quality impacts of different designs. Existing approaches to modeling distributed energy resources address the economic viability and power quality impacts via separate/loosely coupled models. Here, we extend REopt - a techno-economic optimization model developed at National Renewable Energy Laboratory - to consider both within a single model. REopt formulates the design problem as a mixed-integer linear program that solves a deterministic optimization problem for a site's optimal technology mix, sizing, and operation to minimize life cycle cost. REopt has traditionally not constrained the power injection based on power quality. In the work presented here, we expand the REopt platform to consider multiple connected nodes. In order to do this, we model power flow using a fixed-point linear approximation method. Resulting system sizes and voltage magnitudes are validated against the base REopt model, and solutions of established power flow models respectively. We then use the model to explore design considerations of mini-grids in Sub-Saharan Africa. Specifically, we evaluate under what combinations of line length and line capacity it is economically beneficial (or technically required based on voltage limits) to build isolated mini-grids versus an interconnected system that benefits from the economies of scale associated with a single, centralized generation system.

eess.SY