Search NASASearch

SEARCH · Search NASA

Results for “Loading pattern optimization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Experimental evaluation of dynamic data allocation strategies in a distributed database with changing workloads

Traditionally, allocation of data in distributed database management systems has been determined by off-line analysis and optimization. This technique works well for static database access patterns, but is often inadequate for frequently changing workloads. In this paper we address how to dynamically reallocate data for partionable distributed databases with changing access patterns. Rather than complicated and expensive optimization algorithms, a simple heuristic is presented and shown, via an implementation study, to improve system throughput by 30 percent in a local area network based system. Based on artificial wide area network delays, we show that dynamic reallocation can improve system throughput by a factor of two and a half for wide area networks. We also show that individual site load must be taken into consideration when reallocating data, and provide a simple policy that incorporates load in the reallocation decision.

Brunstrom, Anna

An assessment of alternative fuel cell designs for residential and commercial cogeneration

A comparative assessment of three fuel cell systems for application in different buildings and geographic locations is presented. The study was performed at the NASA Lewis Center and comprised the fuel cell design, performance in different conditions, and the economic parameters. Applications in multifamily housing, stores and hospitals were considered, with a load of 10kW-1 MW. Designs were traced through system sizing, simulation/evaluation, and reliability analysis, and a computer simulation based on a fourth-order representation of a generalized system was performed. The cells were all phosphoric acid type cells, and were found to be incompatible with gas/electric systems and more favorable economically than the gas/electric systems in hospital uses. The methodology used provided an optimized energy-use pattern and minimized back-up system turn-on.

Wakefield, R. A.

Time Matters: A Survival Analysis of Public Electric Vehicle Charging Infrastructure Utilization

The rapid adoption of plug-in electric vehicles (PEVs) places significant demands on public charging infrastructure, making it critical to understand and optimize charger utilization. This study provides one of the most comprehensive analyses of charging behavior to date by applying a survival analysis to a dataset of nearly 16 million level 2 (L2) and direct current (DC) fast charger sessions across the United States from 2017 to 2022. Using Kaplan-Meier curves and log rank tests, our analysis reveals statistically significant and distinct duration patterns influenced by charger type, time of day, and day of the week. We find that L2 charging sessions exhibit high variability tied to venue type, whereas DC sessions are more uniform, typically lasting 30-45 min. This study introduces the operational efficiency score (OES), a metric for standardizing the performance evaluation of charging stations. Our findings offer actionable insights for optimizing charger deployment, developing dynamic pricing strategies to reduce vehicle dwell time, and improving load management for grid operators, ultimately enhancing the efficiency and availability of public charging infrastructure.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

The Effects of Partial Mechanical Loading and Ibandronate on Skeletal Tissues in the Adult Rat Hindquarter Suspension Model for Microgravity

We report initial data from a suspended rat model that quantitatively relates chronic partial weightbearing to bone loss. Chronic partial weightbearing is our simulation of the effect of limited artificial gravity aboard spacecraft or reduced planetary gravity. Preliminary analysis of bone by PQCT, histomorphometry, mechanical testing and biochemistry suggest that chronic exposure to half of Earth gravity is insufficient to prevent severe bone loss. The effect of episodic full weightbearing activity (Earth Gravity) on rats otherwise at 50% weightbearing was also explored. This has similarity to treatment by an Earth G-rated centrifuge on a spacecraft that normally maintained artificial gravity at half of Earth G. Our preliminary evidence, using the above techniques to analyze bone, indicate that 2 hours daily of full weightbearing was insufficient to prevent the bone loss observed in 50% weightbearing animals. The effectiveness of partial weightbearing and episodic full weightbearing as potential countermeasures to bone loss in spaceflight was compared with treatment by ibandronate. Ibandronate, a long-acting potent bisphosphonate proved more effective in preventing bone loss and associated functionality based upon structure than our first efforts at mechanical countermeasures. The effectiveness of ibandronate was notable by each of the testing methods we used to study bone from gross structure and strength to tissue and biochemistry. These results appear to be independent of generalized systemic stress imposed by the suspension paradigm. Preliminary evidence does not suggest that blood levels of vitamin D were affected by our countermeasures. Despite the modest theraputic benefit of mechanical countermeasures of partial weightbearing and episodic full weightbearing, we know that some appropriate mechanical signal maintains bone mass in Earth gravity. Moreover, the only mechanism that correctly assigns bone mass and strength to oppose regionally specific force applied to bone is mechanical, a process based upon bone strain. Substantial evidence indicates that the specifics of dynamic loading i.e. time-varying forces are critical. Bone strain history is a predictor of the effect that mechanical conditions have on bone structure mass and strength. Using servo-controlled force plates on suspended rats with implanted strain gauges we manipulated impact forces of ambulation in the frequency (Fourier) domain. Our results indicate that high frequency components of impact forces are particularly potent in producing bone strain independent of the magnitude of the peak force or peak energy applied to the leg. Because a servo-system responds to forces produced by the rat's own muscle activity during ambulation, the direction of ground-reaction loads act on bone through the rat's own musculature. This is in distinction to passive vibration of the floor where forces reach bone through the natural filters of soft tissue and joints. Passive vibration may also be effective, but it may or may not increase bone in the appropriate architectural pattern to oppose the forces of normal ambulatory activity. Effectiveness of high frequency mechanical stimulation in producing regional (muscle directed) bone response will be limited by 1. the sensitivity of bone to a particular range of frequencies and 2. the inertia of the muscles, limiting their response to external forces by increasing tension along insertions. We have begun mathematical modeling of normal ambulatory activity. Effectiveness of high frequency mechanical stimulation in producing regional (muscle directed) bone response will be limited by 1. the sensitivity of bone to a particular range of frequencies and 2. the inertia of the muscles, limiting their response to external forces by increasing tension along insertions. We have begun mathematical modeling of the rat forelimb as a transfer function between impact force and bone strain to predict optimal dynamic loading conditions for this system. We plan additional studies of mechanical counter-measures that incorporate improved dynamic loading, features relevant to anticipated evaluation of artificial gravity, exercise regimens and exposure to Martian gravity, The combination of mechanical countermeasures with ibandronate will also be investigated for signs of synergy.

Schultheis, Lester W.

The Gene Fitness Atlas: A Roadmap for Predicting Evolution

We developed a novel, high-throughput microfluidic device design containing “interaction zones” where progeny cell lines compete against each other allowing for accurate analysis of bacterial cell fitness. The goal of the project was to use the device for two applications: 1) gene knockout screening and 2) antibiotic resistance screening. The microfluidic platform was fabricated using photolithography and soft lithography in polydimethylsiloxane (PDMS). E.coli Keio mutants and fluorescent wildtype parent were chosen for the study. Cells were grown overnight and their loading into the devices and seeding in mother machines was optimized. For mutant screening, the least fit mutant and wildtype parent were cultured individually and then added to the microfluidic device. The mother machines which were seeded with mutant and wildtype were imaged through time lapse microscopy and the growth of cells was observed. For antibiotic screening, wildtype E.coli cells which were grown overnight were added to the device and washed with media containing the antibiotic ampicillin. The growth pattern in presence and absence of ampicillin was observed through time lapse microscopy. It was observed that over a period of four hours, both the mutant and the wildtype divided in the mother machine and pushed daughter cells out into the interaction zone. In case of the antibiotic screening experiment, the fluorescent wildtype divided both in the absence and presence of sublethal concentration of ampicillin. This study is a proof of concept demonstration of high- throughput single cell analysis of cells using a novel microfluidics device.

59 BASIC BIOLOGICAL SCIENCES

Runway Configuration Management with Offline Reinforcement Learning

Runway configuration management (RCM) is a challenging task, and it affects the efficiency of the National Airspace System (NAS) and airport surface operations significantly. Each airport, depending on the geometry, capacity, local climate patterns, etc. has multiple configurations for the runway usage for arriving and departing flights. Many factors such as the incoming/outgoing traffic load, wind direction and speed, convective weather, cloud ceiling and other environmental factors might affect the choice of a runway configuration at any point in time. However, other factors such as safety measures and regulations, noise abatement, capacity of each configuration, and preference of the air traffic controllers (ATCs) can also play a significant role in selecting the configuration. A sub-optimal selection of the runway configuration, or delay in making configuration changes might result in significant increase in taxi times for aircraft on the surface of the airport, fuel and energy use of the aircraft, and maintenance costs. It can also lead to safety concerns, such as an aircraft performing one or more go-arounds before being able to land. All these factors make RCM an extremely important and challenging decision-making process for the ATCs. The current state of practice sets the runway configuration by the ATCs based on relevant information available at the time including weather, traffic, noise abatement, safety bounds, etc. This makes the decision-making process subjective based on the accuracy of the available information and the bias in human decision making. Unfortunately, this approach yields poor results (e.g., significant delays) if the predicted outcomes are uncertain and their relative impact is not well understood. This is especially evident when the uncertainty increases the size of possible predicted outcomes (combinatorial explosion in possible scenarios) that cannot be handled by human reasoning. On the other hand, an automated approach based on machine intelligence can make use of historical data and search through all (or significant amount of) possible scenarios under uncertainty and make well-informed decisions.

Milad Memarzadeh

Overview of Wendelstein 7-X high-performance operation

The Wendelstein 7-X (W7-X) stellarator has completed two consecutive experimental campaigns OP 2.2 (Sep.-Dec. 2024) and OP 2.3 (Feb.-May 2025) under a new operational strategy enabling more than one year of uninterrupted device availability. This approach, supported by exceptionally high subsystem reliability, allowed sustained high-efficiency plasma operations with up to 80–100 discharges per day across a broad range of magnetic configurations. Several key technical upgrades-most notably the first operation of a 1.5 MW class steady-state gyrotron, a new steady-state pellet injector, and advanced real-time feedback control systems significantly enhanced heating, fueling, and plasma control capabilities. Together, these improvements enabled major advances in long-pulse performance, high-β operation, and confinement optimization. Long-pulse discharges achieved 1.8 GJ of injected energy under fully detached divertor conditions, while reduced-field scenarios facilitated record volume-averaged β values approaching 3%. High-performance plasmas with centrally peaked density profiles, created via neutral beam injection (NBI) or sustained pellet fueling, demonstrated strongly reduced turbulent transport and stellarator-record fusion triple products. Complementary studies of power exhaust and divertor heat loads revealed the role of scrape-off-layer drift physics in shaping strike-line patterns under attached conditions. Together, the results from OP 2.2 and OP 2.3 significantly expand the operational space of W7-X and strengthen its role as a leading platform for steady-state stellarator research and reactor-relevant plasma scenarios.

Wendelstein 7-X

Optimal Control Surface Layout for an Aeroservoelastic Wingbox

This paper demonstrates a technique for locating the optimal control surface layout of an aeroservoelastic Common Research Model wingbox, in the context of maneuver load alleviation and active utter suppression. The combinatorial actuator layout design is solved using ideas borrowed from topology optimization, where the effectiveness of a given control surface is tied to a layout design variable, which varies from zero (the actuator is removed) to one (the actuator is retained). These layout design variables are optimized concurrently with a large number of structural wingbox sizing variables and control surface actuation variables, in order to minimize the sum of structural weight and actuator weight. Results are presented that demonstrate interdependencies between structural sizing patterns and optimal control surface layouts, for both static and dynamic aeroelastic physics.

Stanford, Bret K.

Reliability Sensitivity Analysis and Design Optimization of Composite Structures Based on Response Surface Methodology

This report discusses the development and application of two alternative strategies in the form of global and sequential local response surface (RS) techniques for the solution of reliability-based optimization (RBO) problems. The problem of a thin-walled composite circular cylinder under axial buckling instability is used as a demonstrative example. In this case, the global technique uses a single second-order RS model to estimate the axial buckling load over the entire feasible design space (FDS) whereas the local technique uses multiple first-order RS models with each applied to a small subregion of FDS. Alternative methods for the calculation of unknown coefficients in each RS model are explored prior to the solution of the optimization problem. The example RBO problem is formulated as a function of 23 uncorrelated random variables that include material properties, thickness and orientation angle of each ply, cylinder diameter and length, as well as the applied load. The mean values of the 8 ply thicknesses are treated as independent design variables. While the coefficients of variation of all random variables are held fixed, the standard deviations of ply thicknesses can vary during the optimization process as a result of changes in the design variables. The structural reliability analysis is based on the first-order reliability method with reliability index treated as the design constraint. In addition to the probabilistic sensitivity analysis of reliability index, the results of the RBO problem are presented for different combinations of cylinder length and diameter and laminate ply patterns. The two strategies are found to produce similar results in terms of accuracy with the sequential local RS technique having a considerably better computational efficiency.

Rais-Rohani, Masoud

Understanding the structural mechanics of ligated DNA crystals via molecular dynamics simulation

DNA self-assembly is a highly programmable method to construct arbitrary architectures based on sequence complementarity. Among various constructs, DNA crystals are macroscopic crystalline materials formed by assembling motifs via sticky end association. Due to their high structural integrity and size ranging from tens to hundreds of micrometers, DNA crystals offer unique opportunities to study the structural properties and deformation behaviors of DNA assemblies. For example, enzymatic ligation of sticky ends can selectively seal nicks resulting in more robust structures with enhanced mechanical properties. However, the research efforts have been mostly on experiments involving different motif designs, structural optimization, or new synthesis methods, while their mechanics are not yet fully understood. The complex properties of DNA crystals are difficult to study via experiments alone, and numerical simulation can complement and aid the experiments. The coarse-grained molecular dynamics (MD) simulation is a powerful tool that can probe the mechanics of DNA assemblies. Here, we investigate DNA crystals made of four different motif lengths with various ligation patterns (full ligation, major directions, connectors, and in-plane) using oxDNA, an open-source, coarse-grained MD platform. We found that several distinct deformation stages emerge in response to mechanical loading and that the number and the location of ligated nucleotides can significantly modulate structural behaviors. These findings should be useful for predicting crystal properties and thus improving the design.

DNA crystal

Data-Driven Modeling of High-Resolution Residential Load Profiles Using Low-Resolution Smart Meter Measurements

Accurate and high-resolution residential load profiles are essential for power system modeling, demand response planning, and effective grid operation. As the energy sector moves towards a more actively managed distribution system, the ability to understand residential energy consumption at a minute-by-minute scale becomes increasingly critical. High-resolution load profiles provide key insights into demand patterns and user behavior, enabling grid operators to design more effective energy solutions; however, residential load measurements in the field are typically recorded at low resolutions, such as 15-60 minutes, which makes it hard to study the characteristics of different residential customers. This paper addresses these challenges by introducing a data-driven approach to generate realistic, high-resolution residential load profiles based on lowre-solution measurements and weather information. The proposed method retains the key features of the actual residential load measurements while offering appliance-level energy consumption details for each residential building. The results demonstrate the effectiveness of the proposed load profile generator, proving its capability to support utilities in optimizing residential energy management and ensuring a more reliable and resilient grid.

24 POWER TRANSMISSION AND DISTRIBUTION

SWIRP (Submm-Wave and Long Wave InfraRed Polarimeter); Development and Characterization of a Sub-Mm Polarimeter for Ice Cloud Investigations

A major source of uncertainty in climate models is the presence, shape and distribution of ice particles in the uppermost layers of the clouds. The effects of this component are poorly constrained, turning ice particles into an almost-free variable in many climate models.NASA-GSFC is developing a new instrument aimed at measuring the size and shape of ice particles. The instrument consists of two sub-mm polarimeters (at 220 and 670 GHz) coupled with a long-wave infrared polarimeter at 10 micron. Each polarimeter has identical V-pol and H-pol channels; the axes of polarization are defined geometrically by the orientation of the waveguide elements, and the purity has been measured in the lab. The instrument is configured as a conical scanner, suitable for deployment as a payload on a small satellite or on a high-altitude sub-orbital platform. From a 400 km orbit, the instrument has a 3dB spatial resolution of 20 (10) km at 220 (670) GHz and a swath of 600 km over 180 degrees of view.The BAPTA (Bearing And Power Transfer Assembly) carries heritage from the SSMIS design, now in its 22nd year of on-orbit operation, but with a much reduced SWaP (Size Weight and Power) footprint, suitable for a small satellite.The main components of the instrument have been fabricated and are undergoing final testing prior to their integration as a single unit. The sub-mm channels have dedicated secondary reflectors which illuminate a shared primary reflector. The receiving units are placed behind the focal point of the optical arrangement, so that all beams equally illuminate the primary reflector and are almost co-located on the ground (within a single 220 GHz footprint). Primary and secondary beam patterns have been measured and verified to match the as-designed expectations. A Zytex (TM) window is deployed to protect the secondary reflectors and the feed horns from debris and other contaminants, and to reduce the heat load from the active (hot) IR calibration unit. The insertion loss of Zytex has been measured and is accounted in the calibration equation of the sub-mm channels.The radiometric performance of the sub-mm receivers has been characterized in the lab and under operational conditions of temperature and pressure.This paper discusses the design constraints on the sub-mm components, details of the scientific goals and their flowdown, and describes the characterization of the polarimeters. Options to optimize the layout and distribution of the masses within the assembly, with the goal of making the instrument even more compact and fully-compatible with cubesat-class satellites will be presented.

De Amici, G.

Optimization of a Lunar Pallet Lander Reinforcement Structure Using a Genetic Algorithm

In this paper, a unique system level spacecraft design optimization will be presented. A Genetic Algorithm is used to design the global pattern of the reinforcing structure, while a gradient routine is used to adequately stiffen the sub-structure. The system level structural design includes determining the optimal physical location (and number) of reinforcing beams of a lunar pallet lander deck structure. Design of the substructure includes determining placement of secondary stiffeners and the number of rivets required for assembly.. In this optimization, several considerations are taken into account. The primary objective was to raise the primary natural frequencies of the structure such that the Pallet Lander primary structure does not significantly couple with the launch vehicle. A secondary objective is to determine how to properly stiffen the reinforcing beams so that the beam web resists the shear buckling load imparted by the spacecraft components mounted to the pallet lander deck during launch and landing. A third objective is that the calculated stress does not exceed the allowable strength of the material. These design requirements must be met while, minimizing the overall mass of the spacecraft. The final paper will discuss how the optimization was implemented as well as the results. While driven by optimization algorithms, the primary purpose of this effort was to demonstrate the capability of genetic algorithms to enable design automation in the preliminary design cycle. By developing a routine that can automatically generate designs through the use of Finite Element Analysis, considerable design efficiencies, both in time and overall product, can be obtained over more traditional brute force design methods.

Burt, Adam

Investigation of a Cross-Correlation Based Optical Strain Measurement Technique for Detecting radial Growth on a Rotating Disk

The Aeronautical Sciences Project under NASA`s Fundamental Aeronautics Program is extremely interested in the development of novel measurement technologies, such as optical surface measurements in the internal parts of a flow path, for in situ health monitoring of gas turbine engines. In situ health monitoring has the potential to detect flaws, i.e. cracks in key components, such as engine turbine disks, before the flaws lead to catastrophic failure. In the present study, a cross-correlation imaging technique is investigated in a proof-of-concept study as a possible optical technique to measure the radial growth and strain field on an already cracked sub-scale turbine engine disk under loaded conditions in the NASA Glenn Research Center`s High Precision Rotordynamics Laboratory. The optical strain measurement technique under investigation offers potential fault detection using an applied high-contrast random speckle pattern and imaging the pattern under unloaded and loaded conditions with a CCD camera. Spinning the cracked disk at high speeds induces an external load, resulting in a radial growth of the disk of approximately 50.0-im in the flawed region and hence, a localized strain field. When imaging the cracked disk under static conditions, the disk will be undistorted; however, during rotation the cracked region will grow radially, thus causing the applied particle pattern to be .shifted`. The resulting particle displacements between the two images will then be measured using the two-dimensional cross-correlation algorithms implemented in standard Particle Image Velocimetry (PIV) software to track the disk growth, which facilitates calculation of the localized strain field. In order to develop and validate this optical strain measurement technique an initial proof-of-concept experiment is carried out in a controlled environment. Using PIV optimization principles and guidelines, three potential speckle patterns, for future use on the rotating disk, are developed and investigated in the controlled experiment. A range of known shifts are induced on the patterns; reference and data images are acquired before and after the induced shift, respectively, and the images are processed using the cross-correlation algorithms in order to determine the particle displacements. The effectiveness of each pattern at resolving the known shift is evaluated and discussed in order to choose the most suitable pattern to be implemented onto a rotating disk in the Rotordynamics Lab. Although testing on the rotating disk has not yet been performed, the driving principles behind the development of the present optical technique are based upon critical aspects of the future experiment, such as the amount of expected radial growth, disk analysis, and experimental design and are therefore addressed in the paper.

Clem, Michelle M.

Continued Water-Based Phase Change Material Heat Exchanger Development

In a cyclical heat load environment such as low Lunar orbit, a spacecraft's radiators are not sized to reject the full heat load requirement. Traditionally, a supplemental heat rejection device (SHReD) such as an evaporator or sublimator is used to act as a "topper" to meet the additional heat rejection demands. Utilizing a Phase Change Material (PCM) heat exchanger (HX) as a SHReD provides an attractive alternative to evaporators and sublimators as PCM HXs do not use a consumable, thereby leading to reduced launch mass and volume requirements. In continued pursuit of water PCM HX development two full‐scale, Orion sized water‐based PCM HX's were constructed by Mezzo Technologies. These HX's were designed by applying prior research and experimentation to the full scale design. Design options considered included bladder restraint and clamping mechanisms, bladder manufacturing, tube patterns, fill/drain methods, manifold dimensions, weight optimization, and midplate designs. Design and construction of these HX's led to successful testing of both PCM HX's.

Hansen, Scott

Data Structure and Parallel Decomposition Considerations on a Fibonacci Grid

The Fibonacci grid, proposed by Swinbank and Purser (see companion abstract), provides attractive properties for global numerical atmospheric prediction by offering an optimally homogeneous, geometrically regular, and approximately isotropic discretization, with only the polar regions requiring special numerical treatment. It is a mathematical idealization, applied to the sphere, of the multi-spiral patterns often found in botanical structures, such as in pine cones and sunflower heads. Computationally, it is natural to organize the domain, into zones, in each of which the same pair, or triple, of "Fibonacci spirals" dominate. But the further subdivision of such zones into "tiles" of a shape and size suitable for distribution to the processors of a massively parallel computer requires very careful consideration if the subsequent spatial computations along the respective spirals, especially those computations (such as compact differencing schemes) that involve recursion, can be implemented in an efficient "load-balanced "manner without requiring excessive amounts of inter-processor communications. In this paper we show how certain "number theoretic" properties of the Fibonacci sequence (whose numbers prescribe the multiplicity of successive spirals) may be exploited in the decomposition of grid zones into tidy arrangements of triangular grid tiles, each tile possessing one side approximately parallel to the constant-latitude zone boundary. We also describe how the spatially recursive processes may be decomposed across such a tiling, and the directionality of the recursions reversed on alternate grid lines, to ensure a very high degree of load balancing throughout the execution of the computations required for one time step of a global model.

Michalakes, John

Progress of a Cross-correlation Based Optical Strain Measurement Technique for Detecting Radial Growth on a Rotating Disk

The Aeronautical Sciences Project under NASAs Fundamental Aeronautics Program is extremely interested in the development of fault detection technologies, such as optical surface measurements in the internal parts of a flow path, for in situ health monitoring of gas turbine engines. In situ health monitoring has the potential to detect flaws, i.e. cracks in key components, such as engine turbine disks, before the flaws lead to catastrophic failure. In the present study, a cross-correlation imaging technique is investigated in a proof-of-concept study as a possible optical technique to measure the radial growth and strain field on an already cracked sub-scale turbine engine disk under loaded conditions in the NASA Glenn Research Centers High Precision Rotordynamics Laboratory. The optical strain measurement technique under investigation offers potential fault detection using an applied background consisting of a high-contrast random speckle pattern and imaging the background under unloaded and loaded conditions with a CCD camera. Spinning the cracked disk at high speeds induces an external load, resulting in a radial growth of the disk of approximately 50.8-m in the flawed region and hence, a localized strain field. When imaging the cracked disk under static conditions, the disk will appear shifted. The resulting background displacements between the two images will then be measured using the two-dimensional cross-correlation algorithms implemented in standard Particle Image Velocimetry (PIV) software to track the disk growth, which facilitates calculation of the localized strain field. In order to develop and validate this optical strain measurement technique an initial proof-of-concept experiment is carried out in a controlled environment. Using PIV optimization principles and guidelines, three potential backgrounds, for future use on the rotating disk, are developed and investigated in the controlled experiment. A range of known shifts are induced on the backgrounds; reference and data images are acquired before and after the induced shift, respectively, and the images are processed using the cross- correlation algorithms in order to determine the background displacements. The effectiveness of each background at resolving the known shift is evaluated and discussed in order to choose to the most suitable background to be implemented onto a rotating disk in the Rotordynamics Lab. Although testing on the rotating disk has not yet been performed, the driving principles behind the development of the present optical technique are based upon critical aspects of the future experiment, such as the amount of expected radial growth, disk analysis, and experimental design and are therefore addressed in the paper.

Clem, Michelle M.

FPGA-accelerated SpeckleNN with SNL for real-time X-ray single-particle imaging

We present the implementation of a specialized version of our previously published unified embedding model, SpeckleNN, for real-time speckle pattern classification in X-ray Single-Particle Imaging (SPI), using the SLAC Neural Network Library (SNL) on an FPGA platform. This hardware realization transitions SpeckleNN from a prototypic model into a practical edge solution, optimized for running inference near the detector in high-throughput X-ray free-electron laser (XFEL) facilities, such as those found at the Linac Coherent Light Source (LCLS). To address the resource constraints inherent in FPGAs, we developed a more specialized version of SpeckleNN. The original model, which was designed for broader classification across multiple biological samples, comprised ~5.6 million parameters. The new implementation, while reducing the parameter count to 64.6K (a 98.8% reduction), focuses on maintaining the model's essential functionality for real-time operation, achieving an accuracy of 90%. Furthermore, we compressed the latent space from 128 to 50 dimensions. This implementation was demonstrated on the KCU1500 FPGA board, utilizing 71% of available DSPs, 75% of LUTs, and 48% of FFs, with an average power consumption of 9.4W according to the Vivado post-implementation report. The FPGA performed inference on a single image with a latency of 45.015 microseconds at a 200 MHz clock rate. In comparison, running the same inference on an NVIDIA A100 GPU resulted in an average power consumption of ~73W and an image processing latency of around 400 microseconds. Our FPGA-accelerated version of SpeckleNN demonstrated significant improvements, achieving an 8.9 × speedup and a 7.8 × reduction in power consumption compared to the GPU implementation. Key advancements include model specialization and dynamic weight loading through SNL, which eliminates the need for time-consuming FPGA design re-synthesis, allowing fast and continuous deployment of models (re)trained online. These innovations enable real-time adaptive classification and efficient vetoing of speckle patterns, making SpeckleNN more suited for deployment in XFEL facilities. This implementation has the potential to significantly accelerate SPI experiments and enhance adaptability to evolving experimental conditions.

47 OTHER INSTRUMENTATION