Search NASA⌕ Search

SEARCH · Search NASA

Results for “DATA INTEGRITY”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

DMFS: A Data Migration File System for NetBSD

I have recently developed dmfs, a Data Migration File System, for NetBSD. This file system is based on the overlay file system, which is discussed in a separate paper, and provides kernel support for the data migration system being developed by my research group here at NASA/Ames. The file system utilizes an underlying file store to provide the file backing, and coordinates user and system access to the files. It stores its internal meta data in a flat file, which resides on a separate file system. Our data migration system provides archiving and file migration services. System utilities scan the dmfs file system for recently modified files, and archive them to two separate tape stores. Once a file has been doubly archived, files larger than a specified size will be truncated to that size, potentially freeing up large amounts of the underlying file store. Some sites will choose to retain none of the file (deleting its contents entirely from the file system) while others may choose to retain a portion, for instance a preamble describing the remainder of the file. The dmfs layer coordinates access to the file, retaining user-perceived access and modification times, file size, and restricting access to partially migrated files to the portion actually resident. When a user process attempts to read from the non-resident portion of a file, it is blocked and the dmfs layer sends a request to a system daemon to restore the file. As more of the file becomes resident, the user process is permitted to begin accessing the now-resident portions of the file. For simplicity, our data migration system divides a file into two portions, a resident portion followed by an optional non-resident portion. Also, a file is in one of three states: fully resident, fully resident and archived, and (partially) non-resident and archived. For a file which is only partially resident, any attempt to write or truncate the file, or to read a non-resident portion, will trigger a file restoration. Truncations and writes are blocked until the file is fully restored so that a restoration which only partially succeed does not leave the file in an indeterminate state with portions existing only on tape and other portions only in the disk file system. We chose layered file system technology as it permits us to focus on the data migration functionality, and permits end system administrators to choose the underlying file store technology. We chose the overlay layered file system instead of the null layer for two reasons: first to permit our layer to better preserve meta data integrity and second to prevent even root processes from accessing migrated files. This is achieved as the underlying file store becomes inaccessible once the dmfs layer is mounted. We are quite pleased with how the layered file system has turned out. Of the 45 vnode operations in NetBSD, 20 (forty-four percent) required no intervention by our file layer - they are passed directly to the underlying file store. Of the twenty five we do intercept, nine (such as vop_create()) are intercepted only to ensure meta data integrity. Most of the functionality was concentrated in five operations: vop_read, vop_write, vop_getattr, vop_setattr, and vop_fcntl. The first four are the core operations for controlling access to migrated files and preserving the user experience. vop_fcntl, a call generated for a certain class of fcntl codes, provides the command channel used by privileged user programs to communicate with the dmfs layer.

Studenmund, William↗

Metallic fuel transient fuel-cladding interface liquefaction model assessment platform enabled by integrating BISON with databases

A novel platform has been developed within the BISON fuel performance code to assess models of fuel-cladding interface liquefaction for sodium-cooled fast reactor (SFR) metallic fuels. Here, this platform is crucial because liquefaction at the fuel-cladding interface significantly impacts fuel performance and may compromise fuel pin integrity during transient events. To ensure accurate predictions, the platform integrates data collected during the Integral Fast Reactor (IFR) program, now archived in metallic fuel databases. This integration supports verification and validation (V&V) of the models in BISON. Leveraging the extensive US experience with metallic fuel liquefaction and the collections of preserved legacy data, the platform serves as a powerful tool for evaluating existing models and advancing the development of new ones.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

The Contribution of CEOP Data to the Understanding and Modeling of Monsoon Systems

CEOP has contributed and will continue to provide integrated data sets from diverse platforms for better understanding of the water and energy cycles, and for validaintg models. In this talk, I will show examples of how CEOP has contributed to the formulation of a strategy for the study of the monsoon as a system. The CEOP data concept has led to the development of the CEOP Inter-Monsoon Studies (CIMS), which focuses on the identification of model bias, and improvement of model physics such as the diurnal and annual cycles. A multi-model validation project focusing on diurnal variability of the East Asian monsoon, and using CEOP reference site data, as well as CEOP integrated satellite data is now ongoing. Preliminary studies show that climate models have difficulties in simulating the diurnal signals of total rainfall, rainfall intensity and frequency of occurrence, which have different peak hours, depending on locations. Further more model diurnal cycle of rainfall in monsoon regions tend to lead the observed by about 2-3 hours. These model bias offer insight into lack of, or poor representation of, key components of the convective and stratiform rainfall. The CEOP data also stimulated studies to compare and contrasts monsoon variability in different parts of the world. It was found that seasonal wind reversal, orographic effects, monsoon depressions, meso-scale convective complexes, SST and land surface land influences are common features in all monsoon regions. Strong intraseasonal variability is present in all monsoon regions. While there is a clear demarcation of onset, breaks and withdrawal in the Asian and Australian monsoon region associated with climatological intraseasonal variabillity, it is less clear in the American and Africa monsoon regions. The examination of satellite and reference site data in monsoon has led to preliminary model experiments to study the impact of aerosol on monsoon variability. I will show examples of how the study of the dynamics of aerosol-water cycle interactions in the monsoon region, can be best achieved using the CEOP data and modeling strategy.

Lau, William K. M.↗

The contribution of CEOP data to the understanding and modeling of monsoon systems

CEOP has contributed and will continue to provide integrated data sets from diverse platforms for better understanding of the water and energy cycles, and for validating models. In this talk, I will show examples of how CEOP has contributed to the formulation of a strategy for the study of the monsoon as a system. The CEOP data concept has led to the development of the CEOP Inter-Monsoon Studies (CIMS), which focuses on the identification of model bias, and improvement of model physics such as the diurnal and annual cycles. A multi-model validation project focusing on diurnal variability of the East Asian monsoon, and using CEOP reference site data, as well as CEOP integrated satellite data is now ongoing. Similar validation projects in other monsoon regions are being started. Preliminary studies show that climate models have difficulties in simulating the diurnal signals of total rainfall, rainfall intensity and frequency of occurrence, which have different peak hours, depending on locations. Further more model diurnal cycle of rainfall in monsoon regions tend to lead the observed by about 2-3 hours. These model bias offer insight into lack of, or poor representation of key components of the convective,and stratiform rainfall. The CEOP data also stimulated studies to compare and contrasts monsoon variability in different parts of the world. It was found that seasonal wind reversal, orographic effects, monsoon depressions, meso-scale convective complexes, SST and land surface land influences are common features in all monsoon regions. Strong intraseasonal variability is present in all monsoon regions. While there is a clear demarcation of onset, breaks and withdrawal in the Asian and Australian monsoon region associated with climatological intraseasonal variability, it is less clear in the American and Africa monsoon regions. The examination of satellite and reference site data in monsoon has led to preliminary model experiments to study the impact of aerosol on monsoon variability. I will show examples of how the study of the dynamics of aerosol-water cycle interactions in the monsoon region, can be best achieved using the CEOP data and modeling strategy.

Lau, William K. M.↗

LinkML: an open data modeling framework

Background Scientific research relies on well-structured, standardized data; however, much of it is stored in formats such as free-text lab notebooks, nonstandardized spreadsheets, or data repositories. This lack of structure challenges interoperability, making data integration, validation, and reuse difficult. Findings LinkML (Linked Data Modeling Language) is an open framework that simplifies the process of authoring, validating, and sharing data. LinkML can describe a range of data structures, from flat, list-based models to complex, interrelated, and normalized models that utilize polymorphism and compound inheritance. It offers an approachable syntax that is not tied to any one technical architecture and can be integrated seamlessly with many existing frameworks. The LinkML syntax provides a standard way to describe schemas, classes, and relationships, allowing modelers to build well-defined, stable, and optionally ontology-aligned data structures. Once defined, LinkML schemas may be imported into other LinkML schemas. These key features make LinkML an accessible platform for interdisciplinary collaboration and a reliable way to define and share data semantics. Conclusions LinkML helps reduce heterogeneity, complexity, and the proliferation of single-use data models while simultaneously enabling compliance with FAIR (Findable, Accessible, Interoperable, and Reusable) data standards. LinkML has seen increasing adoption in various fields, including biology, chemistry, biomedicine, microbiome research, finance, electrical engineering, transportation, and commercial software development. In short, LinkML makes implicit models explicitly computable and allows data to be standardized at their origin. LinkML documentation and code are available at https://linkml.io/.

AI-ready data↗

Lowering the barrier to access information-rich transient kinetic data for machine learning methods

Transient kinetic data contain a wealth of information about intrinsic features of a catalyst as well as the reaction mechanism. Currently, high volume transient data is underutilized, and data science methods could both increase the value of information that can be extracted from this data, integrate experimental with theoretical data sources, and accelerate the pace of catalyst technology advancement. Transient kinetic characterizations with simple probe molecules exhibiting reversible adsorption, irreversible adsorption and bulk-surface diffusion are presented as training components for similar experiments with more complex surface reactions. In conclusion, by increasing the availability and accessibility of transient kinetic data through details of its structure and acquisition, we aim to decrease the barrier for data scientists to apply machine learning methods to this valuable data source.

Catalysis↗

Integrating State Data Assimilation and Innovative Model Parameterization Reduces Simulated Carbon Uptake in the Arctic and Boreal Region

Model representation of carbon uptake and storage is essential for accurate projection of the response of the arctic-boreal zone to a rapidly changing climate. Land model estimates of LAI and aboveground biomass that can have a marked influence on model projections of carbon uptake and storage vary substantially in the arctic and boreal zone, making it challenging to correctly evaluate model estimates of Gross Primary Productivity (GPP). To understand and correct bias of LAI and aboveground biomass in the Community Land Model (CLM), we assimilated the 8-day Moderate Resolution Imaging Spectroradiometer (MODIS) LAI observation and a machine learning product of annual aboveground biomass into CLM using an Ensemble Adjustment Kalman Filter (EAKF) in an experimental region including Alaska and Western Canada. Assimilating LAI and aboveground biomass reduced these model estimates by 58% and 72%, respectively. The change of aboveground biomass was consistent with independent estimates of canopy top height at both regional and site levels. The International Land Model Benchmarking system assessment showed that data assimilation significantly improved CLM's performance in simulating the carbon and hydrological cycles, as well as in representing the functional relationships between LAI and other variables. Here, to further reduce the remaining bias in GPP after LAI bias correction, we re-parameterized CLM to account for low temperature suppression of photosynthesis. The LAI bias corrected model that included the new parameterization showed the best agreement with model benchmarks. Combining data assimilation with model parameterization provides a useful framework to assess photosynthetic processes in LSMs.

58 GEOSCIENCES↗

Grid Parameters and Voltage Estimation Approach Integrating Data-Driven Converter Model

With Measurements of grid voltage and current are essential for the optimal operation of the grid protection and control (P&C) systems. Grid parameters vary through time during the faults and especially in the converter interfaced resources (CIRs) rich power grid, and thus accurate estimation is critical to avoid the mis-operation of the P&C systems. In this paper, a moving horizon estimation (MHE) as an observer is devised and applied to estimate the grid line parameters and grid voltages for protection enhancement. Due to the proprietary and confidentiality of CIRs, the proposed approach uses the black-box model to represent their dynamics. Leveraging the easily accessible measurements of output current from the black-box model of CIR and voltage at the point of common coupling, the proposed method estimates the grid impedance and grid voltage during normal and faulty operating conditions. The performance shows that the optimization-based observer was able to closely observe the accurate states and parameters, which can be utilized by the P&C systems.

Subedi, Sunil↗

Data for "Integrated Green Biorefinery for the Production of Anthocyanins, Fermentable Sugars, and High Pure Lignin from Miscanthus × giganteus "

Miscanthus x giganteus (Mxg) is a promising perennial crop for producing natural colorants, renewable fuels, and bioproducts. However, natural recalcitrance and high pretreatment cost are major barriers to their complete conversion. In this study, a green processing method has been investigated for efficient recovery of natural pigments (anthocyanins), fermentable sugars, and pure lignin from Mxg genotypes using choline chloride-based natural deep eutectic solvents (NADES) systems. Interestingly, choline chloride: lactic acid (ChCl: LA) NADES-processed biomass resulted in 67.8 ± 2.1 μg g−1 of anthocyanins from dry biomass. A maximum of 87.4%–94.1% glucose yield was achieved after enzymatic saccharification. The effective extraction of lignin with high purity with higher β-aryl ether (βO4) bonds from advanced crops is crucial for lignin valorization. Notably, highly pure lignin (≈93.4% ± 1.4%) is achieved after low-temperature NADES pretreatment while retaining lignin’s native structure. 31P nuclear magnetic resonance demonstrated that total phenolics for ChCl: LA-lignin resulted in 1.20 mmol g−1 hydroxyls. The relative monolignol composition of syringyl (S), guaiacyl (G), and p-hydroxyphenyl (H) is 19.0, 65.7, and 14.3%, respectively, as evidenced by heteronuclear single quantum coherence analysis. This study provides a novel approach for obtaining high-purity lignin for catalytic depolymerization for oligomers and bifunctional monoaromatics production and leverages current cellulosic biorefinery technologies.

biomass analytics↗

Continuous integration data-driven platform of industrial-scale subsurface storage for real-time analytics

This project helped address the growing need for efficient and scalable models to support geological carbon and energy storage, which are crucial for achieving net-zero emissions. Traditionally accurate high-fidelity numerical models have been used to simulate relevant storage processes under a handful of processes, however such models are computationally demanding, making uncertainty quantification impractical. Consequently, we first developed a machine learning framework, based on Graph Neural Operators (GNOs), to improving the accuracy of model predictions for a fixed computational budget. We then developed an Ensemble of Improved Neural Operators (ENO), which uses bagging and Monte Carlo dropout techniques, to further improve prediction accuracy. Lastly, we developed the way to explain progressive transfer learning methods to reduce the amount of training data and computational cost of training (i.e., reduce trainable parameters) when using our models for multiple storage sites. Our numerical investigation, which used real-world case studies, demonstrated that our framework can significantly improve the safety and efficiency of geological storage operations, with potential applications in other domains such as geothermal reservoirs and climate modeling.

54 ENVIRONMENTAL SCIENCES↗

Data for "Integrated ab initio modelling of atomic order and magnetic anisotropy for rare-earth-free magnet design: effects of alloying additions in L1 0 FeNi."

We describe an integrated modelling approach to accelerate the search for novel, single-phase, multicomponent materials with high magnetocrystalline anisotropy (MCA). For a given system we predict the nature of atomic ordering, its dependence on the magnetic state, and then proceed to describe the consequent MCA, magnetisation, and magnetic critical temperature (Curie temperature). Crucially, within our modelling framework, the same ab initio description of a material’s electronic structure determines all aspects. We demonstrate this holistic method by studying the effects of alloying additions in FeNi, examining systems with the general stoichiometries Fe 4 Ni 3 X and Fe 3 Ni 4 X, for additives including X = Pt, Pd, Al, and Co. The atomic ordering behaviour predicted on adding these elements, fundamental for determining a material’s MCA, is rich and varied. Equiatomic FeNi has been reported to require ferromagnetic order to establish the tetragonal L1 0 order suited for significant MCA. Our results show that when alloying additions are included in this material, annealing in an applied magnetic field and/or below a material’s Curie temperature may also promote tetragonal order, along with an appreciable effect on the predicted hard magnetic properties.

36 MATERIALS SCIENCE↗