Search NASASearch

SEARCH · Search NASA

Results for “structured text”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Moving And Working On Space Structures

Clawlike device attaches boots to rails. Memorandum presents, in sketches and brief text, concept for boot-toe clip helping astronaut move about outside on structures being built at Space Station. Clip also helps astronaut maintain stable position at worksite. Concept adaptable to underwater work on such structures as offshore oil rigs.

Mclaughlin, Pat B.

Terrestrial laser scanning data (Levels 0 and 1) from Urban Biogeochemistry Pilot Project sites, Knoxville, Tennessee, Jul 2024 - Jul 2025

This data package contains data from terrestrial laser scanning (TLS) at five urban park sites in Knoxville, Tennessee, USA. All parks include open-grown and/or closed-canopy trees and mixed nearby land use. These study sites were established as part of the Urban Biogeochemistry Pilot Project, which has an overall goal of better understanding how hydrobiogeochemical cycling is altered within the human environment. These five sites represent a gradient of urbanization, and were instrumented to understand hydrological and biogeochemical cycling (e.g., soil moisture, soil physical properties and biogeochemistry, tree transpiration, species type). The TLS data archived here were collected to provide detailed, three-dimensional information about forest structure. Specifically, data were collected to allow tree- and stand-level characterization of woody structure and leaf area. TLS scans were placed to capture the area around trees with sap flow sensors, and as much of a 50 m radius area around the meteorological station as possible given site property limits. Derived products will allow upscaling of water content and transpiration data. This data package contains the following data: - High-level files document further details of the campaign and data package: 1_CampaignSummary.csv provides details about the campaign and study site, 2_ScanAreasDetail.csv provides details about each separate scan area (groups of scans post-processed into a single point cloud), 3_TerrestrialLidarSensor.csv provides further technical details about the Riegl VZ-400i TLS sensor, TLS_CSV_dd.csv is a CSV Data Dictionary providing information about the fields in CSV files following the ESS-DIVE CSV File Formatting Guidelines Reporting Format, TLS_flmd.csv is a File Level Metadata file providing information about each file in the data package following the ESS-DIVE File Level Metadata Reporting Format, and README.txt is a text file describing the overall project and file structure. - Level 0 data are the raw data (.PROJ folders) as recorded by the Riegl VZ-400i TLS instrument before scan co-registration and post-processing with the Riegl's proprietary RiSCAN PRO software, which requires a license. - Level 1 data contain post-processed, co-registered data from each scan area. The "PointClouds" folder for each scan area contains a .las file with 1 cm resolution point cloud data exported from RiSCAN PRO. These are the main files likely to be of interest to most users and can be further processed with any software capable of manipulating .las files (e.g. Python, R CloudCompare). The "Project Information" folder contains log files from post-processing in RiSCAN PRO that may be of interest to users who want to see detailed records of post-processing, including all PDF reports generated by RiSCAN PRO. The "ScanPositions" folder contains information about the final position of all TLS scans, after post-processing, in multiple formats. The file ScanPositions_*.csv provides final geo-referenced scan positions, and the file SOP_backup_*.csv can be used in RiSCAN PRO to restore the co-registered scan positions if users wish to re-process raw data (Level 0 .PROJ folders) with RiSCAN PRO software (e.g., subsample to a different resolution, exclude a certain scan position, or apply different filters on reflectance or deviation values) without redoing time-consuming co-registration steps.

54 ENVIRONMENTAL SCIENCES

Terrestrial laser scanning data (Levels 0 and 1) for Pasoh, Malaysia, Sep 2024

This data package contains data from terrestrial laser scanning (TLS) at the Pasoh Forest Reserve, Malaysia. The Pasoh Forest Reserve is a facility of the Forest Research Institute Malaysia, and contains evergreen lowland dipterocarp forest. The Next-Generation Ecosystem Experiments Tropics (NGEE-Tropics) study areas at Pasoh were established to study how different species respond to climatic variation and soil water availability. Two study areas were chosen representing different topography and species. The TLS data archived here were collected to provide detailed, three-dimensional information about forest structure. Specifically, data were collected to allow tree-level characterization of woody structure and leaf area for 12 focal trees with FloraPulse and sap flux sensors, facilitating estimation of woody biomass and leaf area to allow upscaling of water content and transpiration data to the tree-level. Scan positions were not selected to provide consistent data for non-focal trees with the study areas. This data package contains the following data: - High-level files document further details of the campaign and data package: 1_CampaignSummary.csv provides details about the campaign and study site, 2_ScanAreasDetail.csv provides details about each separate scan area (groups of scans post-processed into a single point cloud), 3_TerrestrialLidarSensor.csv provides further technical details about the Riegl VZ-400i TLS sensor, TLS_CSV_dd.csv is a CSV Data Dictionary providing information about the fields in CSV files following the ESS-DIVE CSV File Formatting Guidelines Reporting Format, TLS_flmd.csv is a File Level Metadata file providing information about each file in the data package following the ESS-DIVE File Level Metadata Reporting Format, and README.txt is a text file describing the overall project and file structure. - Level 0 data are the raw data (.PROJ folders) as recorded by the Riegl VZ-400i TLS instrument before scan co-registration and post-processing with the Riegl's proprietary RiSCAN PRO software, which requires a license. - Level 1 data contain post-processed, co-registered data from each scan area. The "PointClouds" folder for each scan area contains a .las file with 1 cm resolution point cloud data exported from RiSCAN PRO. These are the main files likely to be of interest to most users and can be further processed with any software capable of manipulating .las files (e.g. Python, R CloudCompare). The "Project Information" folder contains log files from post-processing in RiSCAN PRO that may be of interest to users who want to see detailed records of post-processing, including all PDF reports generated by RiSCAN PRO. The "ScanPositions" folder contains information about the final position of all TLS scans, after post-processing, in multiple formats. The file ScanPositions_*.csv provides final geo-referenced scan positions, and the file SOP_backup_*.csv can be used in RiSCAN PRO to restore the co-registered scan positions if users wish to re-process raw data (Level 0 .PROJ folders) with RiSCAN PRO software (e.g., subsample to a different resolution, exclude a certain scan position, or apply different filters on reflectance or deviation values) without redoing time-consuming co-registration steps.

54 ENVIRONMENTAL SCIENCES

LinkML: an open data modeling framework

Background Scientific research relies on well-structured, standardized data; however, much of it is stored in formats such as free-text lab notebooks, nonstandardized spreadsheets, or data repositories. This lack of structure challenges interoperability, making data integration, validation, and reuse difficult. Findings LinkML (Linked Data Modeling Language) is an open framework that simplifies the process of authoring, validating, and sharing data. LinkML can describe a range of data structures, from flat, list-based models to complex, interrelated, and normalized models that utilize polymorphism and compound inheritance. It offers an approachable syntax that is not tied to any one technical architecture and can be integrated seamlessly with many existing frameworks. The LinkML syntax provides a standard way to describe schemas, classes, and relationships, allowing modelers to build well-defined, stable, and optionally ontology-aligned data structures. Once defined, LinkML schemas may be imported into other LinkML schemas. These key features make LinkML an accessible platform for interdisciplinary collaboration and a reliable way to define and share data semantics. Conclusions LinkML helps reduce heterogeneity, complexity, and the proliferation of single-use data models while simultaneously enabling compliance with FAIR (Findable, Accessible, Interoperable, and Reusable) data standards. LinkML has seen increasing adoption in various fields, including biology, chemistry, biomedicine, microbiome research, finance, electrical engineering, transportation, and commercial software development. In short, LinkML makes implicit models explicitly computable and allows data to be standardized at their origin. LinkML documentation and code are available at https://linkml.io/.

AI-ready data

Explanation production by expert planners

Although the explanation capability of expert systems is usually listed as one of the distinguishing characteristics of these systems, the explanation facilities of most existing systems are quite primitive. Computer generated explanations are typically produced from canned text or by direct translation of the knowledge structures. Explanations produced in this manner bear little resemblance to those produced by humans for similar tasks. The focus of our research in explanation is the production of justifications for decisions by expert planning systems. An analysis of justifications written by people for planning tasks has been taken as the starting point. The purpose of this analysis is two-fold. First, analysis of the information content of the justifications will provide a basis for deciding what knowledge must be represented if human-like justifications are to be produced. Second, an analysis of the textual organization of the justifications will be used in the development of a mechanism for selecting and organizing the knowledge to be included in a computer-produced explanation. This paper describes a preliminary analysis done of justifications written by people for a planning task. It is clear that these justifications differ significantly from those that would be produced by an expert system by tracing the firing of production rules. The results from the text analysis have been used to develop an augmented phrase structured grammar (APSG) describing the organization of the justifications. The grammar was designed to provide a computationally feasible method for determining textual organization that will allow the necessary information to be communicated in a cohesive manner.

Bridges, Susan

PDF Entity Annotation Tool (PEAT)

While different text mining approaches – including the use of Artificial Intelligence (AI) and other machine based methods - continue to expand at a rapid pace, the tools used by researchers to create the labeled datasets required for training, modeling, and evaluation remain rudimentary. Labeled datasets contain the target attributes the machine is going to learn; for example, training an algorithm to delineate between images of a car or truck would generally require a set of images with a quantitative description of the underlying features of each vehicle type. Development of labeled textual data that can be used to build natural language machine learning models for scientific literature is not currently integrated into existing manual workflows used by domain experts. Published literature is rich with important information, such as different types of embedded text, plots, and tables that can all be used as inputs to train ML/natural language processing (NLP) models, when extracted and prepared in machine readable formats. Currently, both normalized data extraction of use to domain experts and extraction to support development of ML/NLP models are labor intensive and cumbersome manual processes. Automatic extraction of data and information from formats such as PDFs that are optimized for layout and human readability, not machine readability. The PDF (Portable Document Format) Entity Annotation Tool (PEAT) was developed with the goal of allowing users to annotate publications within their current print format, while also allowing those annotations to be captured in a machine-readable format. One of the main issues with traditional annotation tools is that they require transforming the PDF into plain text to facilitate the annotation process. While doing so lessens the technical challenges of annotating data, the user loses all structure and provenance that was inherent in the underlying PDF. Also, textual data extraction from PDFs can be an error prone process. Challenges include identifying sequential blocks of text and a multitude of document formats (multiple columns, font encodings, etc.). As a result of these challenges, using existing tools for development of NLP/ML models directly from PDFs is difficult because the generated outputs are not interoperable. We created a system that allows annotations to be completed on the original PDF document structure, with no plain text extraction. The result is an application that allows for easier and more accurate annotations. In addition, by including a feature that grants the user the ability to easily create a schema, we have developed a system that can be used to annotate text for different domain-centric schemas of relevance to subject matter experts. Different knowledge domains require distinct schemas and annotation tags to support machine learning.

97 MATHEMATICS AND COMPUTING

Cation Disorder of ${\text{Mg}}_{\mathbf{2}}{\text{SiO}}_{\mathbf{4}}$ in Super‐Earth Mantles

Understanding the mineralogy of exoplanets is essential for unraveling their interior structures, dynamics, and evolution. For large super-Earths, the post-post spinel ${\text{Mg}}_{\mathbf{2}}{\text{SiO}}_{\mathbf{4}}$, one of the major mantle phases, may undergo the order-disorder transition (ODT) at high temperatures. However, the ODT phase boundary of ${\text{Mg}}_{\mathbf{2}}{\text{SiO}}_{\mathbf{4}}$ has not been rigorously constrained. Additionally, fundamental thermodynamic properties of the disordered ${\text{Mg}}_{\mathbf{2}}{\text{SiO}}_{\mathbf{4}}$ remain poorly investigated. Here, we develop a unified machine learning potential (MLP) for ${\text{Mg}}_{\mathbf{2}}{\text{SiO}}_{\mathbf{4}}$ of ab initio accuracy under super-Earth mantle conditions. With the efficient MLP, we extensively calculate the free energy of post-post spinel ${\text{Mg}}_{\mathbf{2}}{\text{SiO}}_{\mathbf{4}}$ via the thermodynamic integration method. The results are used to constrain the ODT phase boundary. Furthermore, we report the P-V-T equation of state and Grüneisen parameters for post-post spinel ${\text{Mg}}_{\mathbf{2}}{\text{SiO}}_{\mathbf{4}}$ across various degrees of disorder. These thermodynamic properties are further applied to update the adiabatic thermal profiles and the mass-radius relation of super-Earths.

36 MATERIALS SCIENCE

Leveraging BERT and Network-Based Attention Analysis for Identifying Treatment Milestones in EHRs

This study introduces a sophisticated data-driven framework for analyzing Electronic Health Records (EHRs) using transformer-based models to identify and disentangle overlapping treatment contexts. The framework leverages a preprocessing pipeline that transforms structured procedural codes into semantically enriched descriptive text, enabling the use of attention mechanisms to cluster medical events into treatment milestones—cohesive and distinct components of care processes. The methodology is rigorously validated using synthetic datasets derived from the MIMIC-III database, designed to simulate the heterogeneity and overlapping procedural contexts characteristic of real-world EHR scenarios. Quantitative evaluation highlights the framework’s robustness in disentangling concurrent care pathways, with attention metrics and unsupervised clustering approaches demonstrating the ability to preserve intra-context relationships while distinguishing inter-context dependencies. By addressing challenges inherent in data heterogeneity, this approach provides a foundation for uncovering complex treatment patterns, advancing clinical decision-making, and optimizing resource allocation in diverse healthcare environments.

Kim, Minsu [ORNL] (ORCID:0000000224185535)

Characterization of International Space Station Crew Members' Workload Contributing to Fatigue, Sleep Disruption and Circadian De-synchronization

The focus of this paper is to characterize how the International Space Station (ISS) crewmembers’ workload may be contributing to sleep loss, circadian misalignment and fatigue. Both sleep quantity and subjective sleep quality are reduced in ISS crewmembers (Barger, Flynn-Evans, Kubey, Walsh, Ronda, Wang, Wright, & Czeisler, 2014). Evidence indicates that the use of hypnotic drugs does not appear to promote extended sleep duration. Because sleep is often driven by psychosocial as well as somatic attributes, traditional therapies may only partially moderate the problem for some individuals. Accordingly, searching for additional abatement tactics is a sensible plan. Scientific studies have shown that sleep can be disrupted from work-related stressors. On the ISS, to optimize their time, the crewmembers follow prescribed, ambitious and rigorous schedules with shared deadlines. Here it will be argued that these human capital leveraging techniques may be undermining the astronaut’s sleep, which could negatively impact performance. Along with half of the Earth-bound working population (Paoli & Merllié, 2001), ISS crewmembers may not be adequately recovering from their workload. This paper begins with a characterization of the working conditions of ISS crewmembers, describes the development of rigorous schedules and portrays a typical workday. Terrestrially-based research is compiled to describe how full and partial sleep deprivation affect physical and cognitive performance and how ISS work characteristics may disrupt sleep and subsequent performance. The literature points toward potential solutions to astronaut fatigue that is related to their workload. Finally, throughout the text, evidence is provided from semi-structured interviews, biographies and textual databases to support the argument that astronaut workload is contributing to their sleep loss and fatigue and that research, development and mitigation strategies should focus on enhancing the restorative process.

fatigue

Structure and Dynamics of Coronal Plasma

Brief summaries of the four published papers produced within the present performance period of NASA Grant NAGW-4081 are presented. The full text of the papers are appended to the report. The first paper titled "Coronal Structures Observed in X-rays and H-alpa Structures" was published in the Kofu Symposium proceedings. The study analyzes cool and hot behavior of two x-ray events, a small flare and a surge. It was found that a large H-alpha surge appears in x-rays as a very weak event, while a weak H-alpha feature corresponds to the brightest x-ray emission on the disk at the time of the observation. Calculations of the heating necessary to produce these signatures, and implications for the driving and heating mechanisms of flares vs. surges are presented. The second paper "Differential Magnetic Field Shear in an Active Region" has been published in The Astrophysical Journal. The study compared the three dimensional extrapolation of magnetic fields with the observed coronal structure in an active region. Based on the fit between observed coronal structure throughout the volume of the region and the calculated magnetic field configurations, the authors propose a differential magnetic field shear model for this active region. The decreasing field shear in the outer portions of the AR may indicate a continual relaxation of the magnetic field with time, corresponding to a net transport of helicity outward. The third paper "Difficulties in Observing Coronal Structure" has been published in the journal Solar Physics. This paper discusses the evidence that the temperature and density structure of the corona are far more complicated than had previously been thought. The discussion is based on five studies carried out by the group on coronal plasma properties, showing that any one x-ray instrument does see all of the plasma present in the corona, that hot and cool material may appear to be co-spatial at a given location in the corona, and that simple magnetic field extrapolations provide only a poor fit to the observed structure. The fourth paper "Analysis and Comparison of Loop Structures Imaged with NIXT and Yohkoh/SXT" has been published in Astronomy and Astrophysics. This paper analyzes and compares a variety of coronal loops, deriving loop pressure and emission measure from loop models. They are able to determine the volume filling factor in the corona, which is found to be in the range 0.001 to 0.01 for compact loops, and of order 1 for large structures. The small values suggest highly filamented structures, especially at lower temperatures.

Golub, Leon

Key elements in building a successful photocomposition applications program

TEXTCOMP, a photocomposition applications program developed by the Technical Information and Documentation Division of the Jet Propulsion Laboratory using the Comp-2 language of Alphanumeric Publication Systems, is described. In Comp-2, each unique block of information such as a paragraph, centered heading, or side heading is termed a 'record'. All commands required for typesetting a record are grouped and identified by a record ID. Once the basic ground rules for typesetting are put into effect by a record ID, changes can be made in texts by means of in-text codes designed to access special characters or perform special operations. The approach used in the design of record ID's and in-text codes for TEXTCOMP to produce a code set structured for maximum usability by non-computer oriented personnel is discussed.

Korbuly, D. K.

Experiences with Text Mining Large Collections of Unstructured Systems Development Artifacts at JPL

Often repositories of systems engineering artifacts at NASA's Jet Propulsion Laboratory (JPL) are so large and poorly structured that they have outgrown our capability to effectively manually process their contents to extract useful information. Sophisticated text mining methods and tools seem a quick, low-effort approach to automating our limited manual efforts. Our experiences of exploring such methods mainly in three areas including historical risk analysis, defect identification based on requirements analysis, and over-time analysis of system anomalies at JPL, have shown that obtaining useful results requires substantial unanticipated efforts - from preprocessing the data to transforming the output for practical applications. We have not observed any quick 'wins' or realized benefit from short-term effort avoidance through automation in this area. Surprisingly we have realized a number of unexpected long-term benefits from the process of applying text mining to our repositories. This paper elaborates some of these benefits and our important lessons learned from the process of preparing and applying text mining to large unstructured system artifacts at JPL aiming to benefit future TM applications in similar problem domains and also in hope for being extended to broader areas of applications.

text mining

Historical perspectives on thermostructural research at the NACA Langley Aeronautical Laboratory from 1948 to 1958

Research on structural problems associated with aerodynamic heating, conducted by the National Advisory Committee for Aeronautics (NACA) during its last decade are described. The text of a special presentation given at the NASA Symposium on Computational Aspects of Heat Transfer in Structure is presented. Some early thermostructural research activities using charts is also discussed. The prinicipal message of the paper is that although vehicle oriented research programs speed development of new technology for specific missions, too much effort may be expended on developing technology which is never used because a vehicle is never built. A healthy research program must provide freedom to explore new ideas that have no obvious applications at the time to generate the technology that makes important, unanticipated flight or vehicle opportunities possible.

Heldenfels, R. R.

Linear elastic fracture mechanics primer

This primer is intended to remove the blackbox perception of fracture mechanics computer software by structural engineers. The fundamental concepts of linear elastic fracture mechanics are presented with emphasis on the practical application of fracture mechanics to real problems. Numerous rules of thumb are provided. Recommended texts for additional reading, and a discussion of the significance of fracture mechanics in structural design are given. Griffith's criterion for crack extension, Irwin's elastic stress field near the crack tip, and the influence of small-scale plasticity are discussed. Common stress intensities factor solutions and methods for determining them are included. Fracture toughness and subcritical crack growth are discussed. The application of fracture mechanics to damage tolerance and fracture control is discussed. Several example problems and a practice set of problems are given.

Wilson, Christopher D.

An evaluation of the Interactive Software Invocation System (ISIS) for software development applications

The Interactive Software Invocation System (ISIS), which allows a user to build, modify, control, and process a total flight software system without direct communications with the host computer, is described. This interactive data management system provides the user with a file manager, text editor, a tool invoker, and an Interactive Programming Language (IPL). The basic file design of ISIS is a five level hierarchical structure. The file manager controls this hierarchical file structure and permits the user to create, to save, to access, and to purge pages of information. The text editor is used to manipulate pages of text to be modified and the tool invoker allows the user to communicate with the host computer through a RUN file created by the user. The IPL is based on PASCAL and contains most of the statements found in a high-level programming language. In order to evaluate the effectiveness of the system as applied to a flight project, the collection of software components required to support the Annular Suspension and Pointing System (ASPS) flight project were integrated using ISIS. The ASPS software system and its integration into ISIS is described.

Noland, M. S.

A general graphical user interface for automatic reliability modeling

Reported here is a general Graphical User Interface (GUI) for automatic reliability modeling of Processor Memory Switch (PMS) structures using a Markov model. This GUI is based on a hierarchy of windows. One window has graphical editing capabilities for specifying the system's communication structure, hierarchy, reconfiguration capabilities, and requirements. Other windows have field texts, popup menus, and buttons for specifying parameters and selecting actions. An example application of the GUI is given.

Liceaga, Carlos A.

ACME: Advanced Combustion via Microgravity Experiments

From Nov. 2017 to Feb. 2022, the ACME project’s six independent investigations of laminar, non-premixed flames of gaseous fuels were conducted on the International Space Station (ISS) in the Combustion Integrated Rack (CIR). An exploration of flames at the extremes of high sooting and high dilution was conducted with a coaxial coflow burner, where the fuel and oxidizer velocities were typically matched. An investigation of electric-field effects also used the same coflow burner as well as a simple gas-jet burner, with a circular electrode mesh downstream of the burner, at voltages of either polarity up to 10 kV. A study focused on material flammability in a quiescent atmosphere emulated the burning of condensed-phased fuels using cylindrical burners with a flat perforated outlet instrumented to measure the heat flux to the burner, i.e., emulated fuel. Three studies of soot processes, flame dynamics, and low-temperature combustion used porous spherical burners yielding a nominally one-dimensional flame structure. The ACME research will be represented through images and text.

flames

WorkJournalMaker (WJMaker) v0.5

The software generates and maintains daily work journal entries in text format, via a web browser. The journal entries are saved in a structured directory file tree on the system running the software. The software also incorporates a database so that it can track the location of files in the file system and various other metadata. The software allows the users to access their journal entries either through the browser or as discrete text files, facilitating sharing and open science. Additionally, to assist with the yearly PMP process, this tool connects to LLM APIs to provide summarization of the journal entries on a month-by-month or weekly basis. The advantage over similar technologies such as Apple Notes (extremely popular for notetaking) is that the instant software does not force the user to stay inside the Apple ecosystem, since it allows for export of the user's text files. This facilitates open science, so that researchers who use the tool can easily transfer their research notes to any other system. The WebJournalMaker repository is here: https://github.com/lbnl-science-it/WorkJournalMaker The WebJournalMaker repository is forked from the JournalSummarizer: https://github.com/tyfong-lbl/JournalSummarizer and builds on its code. I wrote the code for both of these software repos, using generative AI.

Fong, Timothy [Lawrence Berkeley National Laborato