Search NASASearch

SEARCH · Search NASA

Results for “serialization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Challenging conventional assumptions in PV: a high-throughput open-air approach to low-cost perovskite module production

Perovskite solar modules (PSMs) offer a promising pathway to low-cost photovoltaics, yet their commercialization is challenged by manufacturing scalability, device uniformity, additive costs, interlayer complexity, and module stability. This study introduces a comprehensive technoeconomic analysis of single junction PSM's and projections for tandem perovskite-Si modules that integrate all materials and manufacturing steps, module performances, projected lifetimes, and manufacturing costs across scales. Here, we highlight an open-air manufacturing approach to fabricate all active layers of serially interconnected PSMs, including electrodes and charge transport layers, enabling high-throughput production without inert or vacuum environments. The analysis reveals two orders of magnitude throughput enhancement and cost reductions of 24% in all-open-air production, escalating to over 60% at 1 GW factory capacity compared to conventional methods. Levelized cost of energy (LCOE) projections for utility-scale installations over 30 years, accounting for module replacement and recycling, demonstrate the potential to achieve the 2030 US target of $0.03 per kWh with realistic 7–11-year PSM lifetimes, outperforming incumbent silicon-based modules. Neither four terminal (4T) nor two terminal (2T) tandem-Si PSMs improve over single junction perovskite or silicon LCOE regardless of higher efficiencies at any modeled lifetime. Addressing PSM technical challenges with a cost-modeling framework guides commercialization efforts and provides a convincing pathway for challenging incumbent Si-based PV.

14 SOLAR ENERGY

Experience with the alpaka performance portability library in the CMS software

ion Library for Parallel Kernel Acceleration) is a header-only C++ library that provides performance portability across different back-ends, abstracting the underlying levels of parallelism. It supports serial and parallel execution on CPUs, and extremely parallel execution on NVIDIA, AMD and Intel GPUs.This contribution will show how alpaka is used in the CMS software to develop and maintain a single code base; to use different toolchains to build the code for each supported back-end, and link them into a single application; to seamlessly select the best backend at runtime, and implement portable reconstruction algorithms that run efficiently on CPUs and GPUs from different vendors. It will describe the validation and deployment of the alpaka-based implementation in the CMS High Level Trigger, and highlight how it achieves near-native performance.

Alawieh, Jaafar [CERN]

Filling data analysis gaps in time-resolved crystallography by machine learning

There is a growing understanding of the structural dynamics of biological molecules fueled by x-ray crystallography experiments. Time-resolved serial femtosecond crystallography (TR-SFX) with x-ray Free Electron Lasers allows the measurement of ultrafast structural changes in proteins. Nevertheless, this technique comes with some limitations. One major challenge is the quality of data from TR-SFX measurements, which often faces issues like data sparsity, partial recording of Bragg reflections, timing errors, and pixel noise. To overcome these difficulties, conventionally, large volumes of data are collected and grouped into a few temporal bins. The data in each bin are then averaged and paired with the mean of their corresponding jittered timestamps. This procedure provides one structure per bin, resulting in a limited number of averaged structures for the entire time interval spanned by the experiment. Therefore, the information on ultrafast structural dynamics at high temporal resolution is lost. This has initiated research for advanced methods of analyzing experimental TR-SFX data beyond the standard binning and averaging method. To address this problem, we use a machine learning algorithm called Nonlinear Laplacian Spectral Analysis (NLSA), which has emerged as a promising technique for studying the dynamics of complex systems. In this work, we demonstrate the power of this algorithm using synthetic x-ray diffraction snapshots from a protein with significant data incompleteness, timing uncertainties, and noise. Our study confirms that NLSA is a suitable approach that effectively mitigates the effects of these artifacts in TR-SFX data and recovers accurate structural dynamics information hidden in such data.

Trujillo, Justin (ORCID:0000000285505360)

A six degrees of freedom femtosecond laser system for fabrication of small-scale mechanical property specimens

Here, we present the details of a novel ultra-short pulsed laser machining workstation that has been employed for high-throughput laser machining of small-scale mechanical property specimens. This system employs a six degrees of freedom hexapod positioning stage capable of macroscopic movements at high positional accuracy. We developed a methodology that uses quantitative image analysis to measure key parameters required to minimize the hexapod positioning and rotational error. Application of this system to laser machining of small-scale 316L stainless steel tensile specimens and ultra-high molecular weight polyethylene compressive specimens using eucentric tilt and rotation about the specimen axis will be shown, where serial laser milling at a specimen tilt angle of 10° was used to effectively eliminate any taper in the sample cross section that is typically found in laser machining.

47 OTHER INSTRUMENTATION

Nm-scale electrical resistance imaging on CdTe by scanning spreading resistance microscopy

Local resistance imaging can provide information on nm-scale carrier distribution in semiconductor devices. Scanning spreading resistance microscopy (SSRM), an atomic force microscopy-based nm-scale resistance mapping technique, has been developed for carrier delineation in Si microdevices. We report on the development and validation of SSRM on CdTe materials, by testing on molecular beam epitaxy (MBE) grown CdTe films. The probe/CdTe contact resistance was suppressed sufficiently below sample's spreading resistance by pressing the probe into the sample with ∼mN contact force and applying a large sample/probe forward bias voltage (V s ), which was understood by analyzing current-voltage (I-V) involving a serially connected insulating top layer with underlying spreading resistance. The carrier concentration as deduced from the resistance measurement, using a single mobility value, is consistent with Hall measurement with a standard deviation of 14% based on a set of MBE films with carrier concentrations in the range of 10 15 –10 16 /cm 3 . The doping polarity was readily identified by flipping V s polarity, where the resistance with reverse V s is orders of magnitude larger than forward V s . While focusing on the SSRM technique validation, we also show an example on an As-doped Cd(Se,Te) polycrystalline thin film of a high-performance CdTe solar cell, which illustrates the local resistance nonuniformity with up to two orders of magnitude differences, indicating if local mobility is roughly constant, local carrier concentration can have significant nonuniformity.

14 SOLAR ENERGY

Metrology for femtosecond pulsed x-ray heating in diamond anvil cell experiments at the European XFEL: Revisiting the iron phase diagram up to 150 GPa

The development of pulsed intense x-ray sources, such as free electron laser, offers new avenues for high pressure experiments. Here, we study the feasibility and metrology of x-ray heating in diamond anvil cells at the European x-ray free electron laser. This method enables one to volumetrically heat the sample while inhibiting chemical migration and probing the crystallographic structure of the sample throughout the heating with a high repetition rate. We focus our study on iron, whose phase diagram is well established up to 100 GPa, to explore the possibilities and limitations of this technique. We volumetrically heat iron samples at starting pressures ranging from 10 to 138 GPa, using the x-ray beam pulsed at 4.5 MHz in a serial pump-and-probe experimental design. Experimental challenges arise from temperature gradients within the sample, changes in temperature at the 100 ns timescale, the difficulty of direct temperature estimates, the effect of thermal pressure, and the presence of metastable crystallites due to rapid cycles of heating and cooling. Hence, we develop a multi-crystal-like data processing method that allows us to account for sample heterogeneity in probed conditions. We then calibrate our measurements using known physical properties of iron under pressure. Thermal pressure in our experiments increases from 4% of the isochoric prediction at 10 GPa to 23% at 138 GPa, and we show that our data are in agreement with most previous observations of iron in this pressure range. The method can now be implemented at higher pressures and temperatures and on materials with unknown phase diagrams.

Materials science

Reuniting crystallography with real space: Ab initio structure elucidation with 4D-STEM

Structure elucidation via single-crystal methods has historically lacked experimental access to real-space information, instead relying exclusively on diffraction-space measurements of Bragg reflections. Here we exploit the dual-space imaging power of 4D scanning transmission electron microscopy to meaningfully integrate real-space information into the crystallographic workflow. We show that virtual apertures assembled by segmentation of high-angle annular dark-field images enable i) pixel-by-pixel separation of coherent Bragg signal from clusters of closely spaced nanocrystals and ii) selective extraction of integrated intensities from thinner subregions of individual specimens, facilitating retroactive tuning of multiple scattering artifacts. This strategy empowers us to simply pick and choose whichever nanoscale regions of interest generate the highest-quality diffraction patterns, allowing us to solve several independent structures of the metal-organic framework UiO-66 from specimens whose agglomerated morphology proved intractable for conventional microcrystal electron diffraction. Our method is compatible with both rotational and serial approaches to data processing, ultimately divulging the first scanning nanobeam electron diffraction structures determined by direct methods at subangstrom resolution.

Saha, Ambarneil

Cryogenic electron tomography by the numbers: Charting underexplored lineages in structural cell biology

Imaging cells and their interactions across the whole biosphere with molecular-scale resolution is key for understanding structure–function relations. Cryogenic electron tomography (cryo-ET) is a powerful method for obtaining this critical information. However, cryo-ET studies are challenging and often limited to a small number of cell types per study. Here, we collate cryo-ET data from hundreds of cells and tissues across the biosphere to i) identify emerging methodological trends, ii) pinpoint strategies to reduce imaging time and costs, iii) quantitatively compare methods for cell freezing and sectioning, and iv) census cryo-ET species coverage across all domains of life. Comparing the fraction of cellular material within a single lamella across all domains of life reveals an order of magnitude difference between eukaryotes (1%) compared to bacteria (9%) and archaea (14%). We calculate the fraction of cellular material which can be imaged using distinct sectioning methods on multicellular communities and tissues—identifying serial lift-out as a powerful approach for obtaining more complete cellular depictions. Finally, we show that the biodiversity of current cryo-ET studies is 2 to 3 orders of magnitude lower than in sequence libraries and 4 to 5 lower than the total predicted on Earth. Our analyses reveal major evolutionary lineages which remain critically understudied and highlight where future cryo-ET research would be most impactful.

HPF

Optimizing Charge-coupled Device Readout Enabled by the Floating-gate Amplifier

Multiple-Amplifier Sensing (MAS) charge-coupled devices (CCDs) have recently been shown to be promising silicon detectors that meet noise sensitivity requirements for next generation Stage-5 spectroscopic surveys and potentially, future space-based imaging of extremely faint objects on missions such as the Habitable Worlds Observatory. Building upon the capability of the Skipper CCD to achieve deeply sub-electron noise floors, MAS CCDs utilize multiple floating-gate amplifiers along the serial register to increase the readout speed by a factor of the number of output nodes compared to a Skipper CCD. We introduce and experimentally demonstrate on a 16-channel prototype device new readout techniques that exploit the MAS CCD’s floating-gate amplifiers to optimize the correlated double sampling by resetting once per line instead of once per pixel. With this new mode, we find an optimal filter to subtract the noise from the signal during read out. We also take advantage of the MAS CCD’s structure to tune the read time by independently changing integration times for the signal and reference level. Together with optimal weighted averaging of the 16 outputs, these approaches enable us to reach a sub-electron noise of 0.9 e − rms pix −1 at 19 μs pix −1 for a single charge measurement per pixel—simultaneously giving a 30% faster readout time and 10% lower read noise compared to performance previously evaluated without these techniques.

Lin, Kenneth W. (ORCID:0000000189672281)

RD53 pixel readout integrated circuits for ATLAS and CMS HL-LHC upgrades

The RD53 collaboration has since 2013 developed new hybrid pixel detector chips with 50 × 50 μm2 pixels for the HL-LHC upgrades of the ATLAS and CMS experiments at CERN. A common architecture, design and verification framework has been developed to enable final pixel chips of different sizes to be designed, verified and tested to handle extreme hit rates of 3 GHz/cm2 (up to 12 GHz per chip) together with an increased trigger rate of 1 MHz and efficient readout of up to 5.12 Gbits/s per pixel chip. Tolerance to an extremely hostile radiation environment with 1 Grad over 10 years and induced SEU (Single Event Upset) rates of up to 100 upsets per second per chip have been major challenges to make reliable pixel chips. Three generations of pixel chips, and many specific mixed signal building blocks and radiation test chips, have been submitted and extensively tested to get to final production chips. The large, complex and high rate pixel chips have been developed with a strong emphasis on low power consumption together with a concurrent development and qualification of novel serial powering at chip, module and system level, to minimize detector material budget.

Alimonti, G

How transparent is graphene? A surface science perspective on remote epitaxy

Remote epitaxy is the synthesis of a single crystalline film on a graphene-covered substrate, where the film adopts epitaxial registry to the substrate as if the graphene is transparent. Despite many exciting applications for flexible electronics, strain engineering, and heterogeneous integration, an understanding of the fundamental synthesis mechanisms remains elusive. Here we offer a perspective on the synthesis mechanisms, focusing on the foundational assumption of graphene transparency. We identify challenges for quantifying the strength of the remote substrate potential that permeates through graphene, and propose Fourier and beating analysis as a bias-free method for decomposing the lattice potential contributions from the substrate, from graphene, and from surface reconstructions, each at different frequencies. We highlight the importance of graphene-induced reconstructions on epitaxial templating, drawing comparison to moiré epitaxy. We highlight the role of the remote potential in tuning surface diffusion and adatom kinetics on graphene, which are crucial for navigating the competition between remote epitaxy and defect-seeded mechanisms like pinhole epitaxy. In light of this weak remote potential, we re-evaluate the current state-of-the-art experimental evidence, highlighting why it remains challenging to experimentally validate a ‘remote’ epitaxy mechanism that cannot be explained by alternatives, such as pinhole-seeded epitaxy or serial van der Waals epitaxy. We end with one experimental example that, to out knowledge, cannot be explained by competing mechanisms: a different long-range epitaxial relationship for GdPtSb films grown on graphene/sapphire, compared to direct epitaxy on sapphire. We suggest for future experiments that directly measure the remote potential and impact of tuneable growth kinetics.

epitaxy

Deciphering the altered conformational states of bifunctional thaumarchaeal crotonyl-CoA hydratase and 3-hydroxypropionyl-CoA dehydratase from Nitrosopumilus maritimus

Abstract The thaumarchaeal 3-hydroxypropionate/4-hydroxybutyrate (3HP/4HB) cycle represents one of the most efficient mechanisms for CO2 fixation discovered to date. Within this cycle, the enzyme encoded by Nmar_1308 from Nitrosopumilus maritimus SCM1 plays a crucial role due to its dual functionality as both a crotonyl-CoA hydratase (CCAH) and a 3-hydroxypropionyl-CoA dehydratase (3HPD). Although the importance of a bifunctional enzyme for lowering the cost of biosynthesis, the details of structural dynamics are still missing. Here, in addition to our cryogenic temperature structures, we determined the first ambient temperature structures of the Nmar_1308 protein by Serial Femtosecond X-ray Crystallography (SFX). The determined structures capture previously unobserved conformational dynamics of the Nmar_1308 protein, providing invaluable information for future synthetic biology applications.

Destan, Ebru (ORCID:0000000231290827)

Integrated System Planning: Emerging Software Requirements in the Power Industry

Power system planning software remains fragmented across organizational boundaries, with specialized tools for capacity expansion, production cost modeling, power flow, and dynamic analysis operating on incompatible data models and assumptions. This article argues that the fragmentation is not merely a technical problem but a predictable consequence of Conway's law: software architectures mirror the departmental structures within which they are developed. Regulatory milestones like Federal Energy Regulatory Commission (FERC) Order 888 formalized these divisions, but the roots trace back to the distinct engineering disciplines-mechanical, chemical, and electrical-that staffed generation and transmission planning departments in vertically integrated utilities. As the industry moves toward integrated system planning (ISP) that coordinates generation, transmission, and distribution investment decisions, the software ecosystem must evolve accordingly. We identify five categories of software requirements to enable this transition: coherent data inputs decoupled from individual applications, unified and extensible data schemas, modular component representations that support multiple abstraction levels, lifecycle management of planning datasets, and well-defined application programming interface (API) contracts that separate data exchange from algorithmic control. We examine how these requirements interact with three common workflow patterns-serial gate clearing, sequential multiapplication, and convergence oriented-and discuss the interface design principles each demands. We then outline a vision for platform-based planning architectures where specialized analytical services compose through standardized interfaces and where artificial intelligence (AI)/machine learning (ML) tools augment decision support within a disciplined software infrastructure. The practices proposed here offer a path from today's siloed tool collections toward collaborative planning ecosystems capable of handling the complexity of modern power system transformation.

24 POWER TRANSMISSION AND DISTRIBUTION

A 32-Channel Cryo-CMOS ASIC for SNSPD Biasing and Readout with Picosecond

Superconducting nanowire single-photon detectors (SNSPD) are a promising technology for particle detection. Although SNSPDs have demonstrated picosecond timing accuracy, scaling up large arrays has proved challenging. In this work, we introduce a 32-channel cryo-CMOS application-specific integrated circuit (ASIC) that can be tightly integrated with SNSPD arrays. The ASIC is designed to operate at a temperature of 4K and can perform up to 32 simultaneous timing measurements with a root-mean-square (RMS) accuracy of 8.0ps. The ASIC includes on-chip circuitry for externally biasing superconducting devices, low-noise amplifiers for reading superconducting devices, high-resolution time-to-digital converters (TDC) for time-tagging events, and serializers for transmitting data to room-temperature electronics. The ASIC is manufactured in a 22nm FDSOI process and occupies an area of 4.0mm x 1.0mm. The performance of the ASIC was verified using custom cryogenic device models internally developed for the 22nm SOI process. Measurement results will be presented at the conference.

Fredenburg, Jeff

A Cryogenic readout integrated circuit with analog pile-up and in-Pixel ADC for high frame rate Skipper CCD-in-CMOS Sensors

The Skipper CCD-in-CMOS Parallel Read-Out Circuit V2 (SPROCKET2) is designed to enable high frame rate readout of Skipper CCD-in-CMOS image sensors. The SPROCKET2 pixel is fabricated in a 65 nm CMOS process and occupies a 60$\mu$m $\times$ 60$\mu$m footprint. SPROCKET2 is intended to be heterogeneously integrated with a pixelated Skipper CCD-in-CMOS sensor, such that one readout pixel is connected to a multiplexed array of 16 active image sensor pixels, to match their spatial geometry. Our design benefits from the Skipper CCD-in-CMOS sensor's non-destructive readout capability to achieve exceptionally low noise through multi-sampling and averaging while optimizing for total power consumption. The pixel readout utilizes correlated double sampling to minimize 1/f noise and includes "pile-up" of ten successive samples in the analog domain before digitizing at a rate of 66.7 ksps. Measurement results of in-pixel serial SAR ADC show DNL and INL of ~0. 44 LSB and 0.58 LBS respectively. A large area array of 20,000 SPROCKET2 ADC pixels (multiplexed 1:16 to 320,000 sensor pixels) is currently under test. By reading out data over a 10 Gbps optical link, this pixel design enables a frame rate of $\sim$ 4 kfps for large sensing areas with minimal sensing deadtime. In the highest gain mode, the pixelated ADC has an input-referred resolution of 10$\mu$V with a simulated power consumption of 50$\mu$W. The pixel operates with constant current draw to minimize power-rail crosstalk.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND

MIND-MAC: Multi-Level In-memory Quasi Non-Destructive MAC Operation in Compact 2T-nC FeRAM for Efficient DNN Accelerator

We present MIND-MAC, a compact 2T-nC FeRAM architecture that performs multi-level, quasi-non-destructive in-memory multiply–accumulate (MAC) for deep neural networks. By exploiting voltage-controlled partial domain switching in MFM capacitors and read-transistor amplification, the cell stores multi-bit weights and gates bit-serial inputs to produce an accumulated current on shared lines. We combine TCAD-extracted parasitics with experimentally calibrated ferroelectric models in SPICE to validate device-/circuit-level behavior, and validate multi-level sensing and QNRO with measurements on a fabricated 2T-3C test vehicle. An analytical system model maps MIND-MAC to a 6-GB main-memory in-memory compute (IMC) architecture and benchmarks VGG13 inference in 61.08 ms at 964.99 mJ. Results indicate high density, reduced rewrite overhead, and energy efficiency, positioning 2T-nC FeRAM as a promising IMC candidate for next-generation AI hardware.

36 MATERIALS SCIENCE

Optimizing Management of Persistent Data Structures in High-Performance Analytics

Large-scale data analytics workflows ingest massive input data into various data structures, including graphs and key-value datastores. These data structures undergo multiple transformations and computations and are typically reused in incremental and iterative analytics workflows. Persisting in-memory views of these data structures enables reusing them beyond the scope of a single program run while avoiding repetitive raw data ingestion overheads. Memory-mapped I/O enables persisting in-memory data structures without data serialization and deserialization overheads. However, memory-mapped I/O lacks the key feature of persisting consistent snapshots of these data structures for incremental ingestion and processing. The obstacles to efficient virtual memory snapshots using memory-mapped I/O include background writebacks outside the application’s control, and the significantly high storage footprint of such snapshots. To address these limitations, we present Privateer, a memory and storage management tool that enables storage-efficient virtual memory snapshotting while also optimizing snapshot I/O performance. Here, we integrated Privateer into Metall, a state-of-the-art persistent memory allocator for C++, and the Lightning Memory-Mapped Database (LMDB), a widely-used key-value datastore in data analytics and machine learning. Privateer optimized application performance by 1.22× when storing data structure snapshots to node-local storage, and up to 16.7× when storing snapshots to a parallel file system. Privateer also optimizes storage efficiency of incremental data structure snapshots by up to 11× using data deduplication and compression.

Computer science

Performance-Aligned LLMs for Generating Fast HPC Code

Optimizing scientific software is a difficult task because codebases are often large and complex, and performance can depend upon several factors including the algorithm, its implementation, and hardware among others. Causes of poor performance can originate from disparate sources and be difficult to diagnose. Recent years have seen a multitude of work that use large language models (LLMs) to assist in software development tasks. However, these tools are trained to model the distribution of code as text, and are not specifically designed to understand performance aspects of code. In this work, we introduce a reinforcement learning based methodology to align the outputs of code LLMs with performance. This allows us to build upon the current code modeling capabilities of LLMs and extend them to generate better performing code. Here, we demonstrate that our fine-tuned model improves the expected speedup of generated code over base models for a set of benchmark tasks from 0.9 to 1.6 for serial code and 1.9 to 4.5 for OpenMP parallel code.

Computer science