Search NASA⌕ Search

SEARCH · Search NASA

Results for “text analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

DESI DR2 Baryon Acoustic Oscillations from the Lyman Alpha Forest Multipoles

We present an alternative measurement of the Baryon Acoustic Oscillation (BAO) using the Legendre multipole representation of the Ly$α$ forest correlation functions from the second data release (DR2) of the Dark Energy Spectroscopic Instrument survey. Compressing the auto- and cross-correlation functions into Legendre multipoles yields a positive-definite covariance matrix without any smoothing -- unlike the baseline DR2 analysis -- thanks to a significantly reduced data vector size. We introduce the statistical corrections required to debias the finite-sample covariance matrix estimate and demonstrate that monopole and quadrupole terms for both auto- and cross-correlations can be used even when the correlation functions are distorted by continuum errors and contaminated by metals. This formalism has slightly diminished the constraining power of the BAO scale, while considerably weakening constraints on nuisance parameters. We measure the isotropic BAO scale with $0.93\%$ precision at $z_\mathrm{eff}=2.35$, the Hubble parameter $H(z_\mathrm{eff})=(239.5\pm3.4)~(147.09~\mathrm{Mpc}/r_d) ~\mathrm{km~s}^{-1}~\text{Mpc}^{-1}$, and the transverse comoving distance $D_M(z_\mathrm{eff})=(5.80 \pm 0.10)~(r_d/147.09~\mathrm{Mpc})$~Gpc for a given value of the sound horizon ($r_d$). Our BAO results are entirely consistent with the baseline DR2 analysis.

Karaçaylı, Naim Göksel [Chicago U., KICP; Ohio Sta↗

A structural equation modeling approach to leveraging the power of extant sentiment analysis tools

Machine-derived sentiment analysis has become a pervasive and useful tool to address a wide array of issues in natural language processing. Leading technology companies such as Google now provide sentiment analysis tools (SATs) as readily accessible online products. Academic researchers develop and make available SATs to support the research enterprise. One of the major challenges with SATs is the inconsistencies in results among the various SATs. Consequently, the selection of a SAT for a specific purpose may significantly impact the application. This study addresses the foregoing problem by utilizing structural equation modeling to merge the outputs of SATs to develop a combined sentiment metric without the need for a labeled training dataset. This method is applicable to a wide range of text-based problems, is data-driven, and replicable. It was tested using three publicly available datasets and compared against seven different SATs. The results indicate that as a continous measure, the proposed method outperformed other SATs in the movie reviews and SemEval datasets, and achieved a tie for first place with IBM Watson on the Sentiment 140 dataset. Also, compared to the published major alternatives, the arithmetic mean solution, this approach performed better across these three datasets.

97 MATHEMATICS AND COMPUTING↗

Dark matter substructure or source model systematics? A case study of cluster lens Abell S1063

Mapping the small-scale structure of the universe through gravitational lensing is a promising tool for probing the particle nature of dark matter. Curved Arc Basis (CAB) has been proposed as a local lensing formalism in galaxy clusters, with the potential to detect low-mass dark matter substructure. In this work, we analyse the cluster lens Abell S1063 in search of dark matter substructure with the CAB formalism, using multiband imaging data from James Webb Space Telescope ( JWST ). We use two different source modelling methods: shapelets and pixel-based source reconstruction based on Delaunay triangulation. We find that source modelling systematics from shapelets result in a disagreement between CAB parameters measured from different filters. Source modelling with Delaunay significantly alleviates this systematic, as seen in the improvement in agreement across filters. We also find that inadequate complexity in source modelling can result in convincing spurious detections of dark matter substructure from strong gravitational lenses, as seen by our $\Delta \text{BIC} > 20$ measurement of a $M \sim 10^{10}$ ${\rm M}_{\odot }$ subhalo with shapelets, a spurious detection that is not reproduced with Delaunay source modelling. We demonstrate that multiband analysis with different JWST filters is key for disentangling source and lens model systematics from dark matter substructure detections.

79 ASTRONOMY AND ASTROPHYSICS↗

ATLAS, an integrated structural analysis and design system. Volume 3: User's manual, input and execution data

The input data and execution control statements for the ATLAS integrated structural analysis and design system are described. It is operational on the Control Data Corporation (CDC) 6600/CYBER computers in a batch mode or in a time-shared mode via interactive graphic or text terminals. ATLAS is a modular system of computer codes with common executive and data base management components. The system provides an extensive set of general-purpose technical programs with analytical capabilities including stiffness, stress, loads, mass, substructuring, strength design, unsteady aerodynamics, vibration, and flutter analyses. The sequence and mode of execution of selected program modules are controlled via a common user-oriented language.

Dreisbach, R. L.↗

The Cool Flames Experiment

A space-based experiment is currently under development to study diffusion-controlled, gas-phase, low temperature oxidation reactions, cool flames and auto-ignition in an unstirred, static reactor. At Earth's gravity (1g), natural convection due to self-heating during the course of slow reaction dominates diffusive transport and produces spatio-temporal variations in the thermal and thus species concentration profiles via the Arrhenius temperature dependence of the reaction rates. Natural convection is important in all terrestrial cool flame and auto-ignition studies, except for select low pressure, highly dilute (small temperature excess) studies in small vessels (i.e., small Rayleigh number). On Earth, natural convection occurs when the Rayleigh number (Ra) exceeds a critical value of approximately 600. Typical values of the Ra, associated with cool flames and auto-ignitions, range from 104-105 (or larger), a regime where both natural convection and conduction heat transport are important. When natural convection occurs, it alters the temperature, hydrodynamic, and species concentration fields, thus generating a multi-dimensional field that is extremely difficult, if not impossible, to be modeled analytically. This point has been emphasized recently by Kagan and co-workers who have shown that explosion limits can shift depending on the characteristic length scale associated with the natural convection. Moreover, natural convection in unstirred reactors is never "sufficiently strong to generate a spatially uniform temperature distribution throughout the reacting gas." Thus, an unstirred, nonisothermal reaction on Earth does not reduce to that generated in a mechanically, well-stirred system. Interestingly, however, thermal ignition theories and thermokinetic models neglect natural convection and assume a heat transfer correlation of the form: q=h(S/V)(T(bar) - Tw) where q is the heat loss per unit volume, h is the heat transfer coefficient, S/V is the surface to volume ratio, and (T(bar) - Tw ) is the spatially averaged temperature excess. This Newtonian form has been validated in spatially-uniform, well-stirred reactors, provided the effective heat transfer coefficient associated with the unsteady process is properly evaluated. Unfortunately, it is not a valid assumption for spatially-nonuniform temperature distributions induced by natural convection in unstirred reactors. "This is why the analysis of such a system is so difficult." Historically, the complexities associated with natural convection were perhaps recognized as early as 1938 when thermal ignition theory was first developed. In the 1955 text "Diffusion and Heat Exchange in Chemical Kinetics", Frank-Kamenetskii recognized that "the purely conductive theory can be applied at sufficiently low pressure and small dimensions of the vessel when the influence of natural convection can be disregarded." This was reiterated by Tyler in 1966 and further emphasized by Barnard and Harwood in 1974. Specifically, they state: "It is generally assumed that heat losses are purely conductive. While this may be valid for certain low pressure slow combustion regimes, it is unlikely to be true for the cool flame and ignition regimes." While this statement is true for terrestrial experiments, the purely conductive heat transport assumption is valid at microgravity (mu-g). Specifically, buoyant complexities are suppressed at mu-g and the reaction-diffusion structure associated with low temperature oxidation reactions, cool flames and auto-ignitions can be studied. Without natural convection, the system is simpler, does not require determination of the effective heat transfer coefficient, and is a testbed for analytic and numerical models that assume pure diffusive transport. In addition, mu-g experiments will provide baseline data that will improve our understanding of the effects of natural convection on Earth.

Pearlman, Howard↗

Statistical properties of DNA sequences

We review evidence supporting the idea that the DNA sequence in genes containing non-coding regions is correlated, and that the correlation is remarkably long range--indeed, nucleotides thousands of base pairs distant are correlated. We do not find such a long-range correlation in the coding regions of the gene. We resolve the problem of the "non-stationarity" feature of the sequence of base pairs by applying a new algorithm called detrended fluctuation analysis (DFA). We address the claim of Voss that there is no difference in the statistical properties of coding and non-coding regions of DNA by systematically applying the DFA algorithm, as well as standard FFT analysis, to every DNA sequence (33301 coding and 29453 non-coding) in the entire GenBank database. Finally, we describe briefly some recent work showing that the non-coding sequences have certain statistical features in common with natural and artificial languages. Specifically, we adapt to DNA the Zipf approach to analyzing linguistic texts. These statistical properties of non-coding sequences support the possibility that non-coding regions of DNA may carry biological information.

Non-NASA Center↗

Helicopter theory

A comprehensive presentation is made of the engineering analysis methods used in the design, development and evaluation of helicopters. After an introduction covering the fundamentals of helicopter rotors, configuration and operation, rotary wing history, and the analytical notation used in the text, the following topics are discussed: (1) vertical flight, including momentum, blade element and vortex theories, induced power, vertical drag and ground effect; (2) forward flight, including in addition to momentum and vortex theory for this mode such phenomena as rotor flapping and its higher harmonics, tip loss and root cutout, compressibility and pitch-flap coupling; (3) hover and forward flight performance assessment; (4) helicopter rotor design; (5) rotary wing aerodynamics; (6) rotary wing structural dynamics, including flutter, flap-lag dynamics ground resonance and vibration and loads; (7) helicopter aeroelasticity; (8) stability and control (flying qualities); (9) stall; and (10) noise.

Johnson, W.↗

Searching for Neutrino Tridents in the NOvA Near Detector

This dissertation presents a search for neutrino trident production in the NOvA near detector through the coherent ``dimuon" channel: $\nu_\mu +\hspace{1pt}\text{X} \rightarrow \nu_\mu + \mu^- + \mu^+ +\hspace{1pt}\text{X}$. Trident production is a rare, purely electroweak process with sensitivity to physics beyond the Standard Model. The theoretical background, motivation for studying the process, and previous experimental measurements are reviewed. The analysis uses data collected by the NOvA near detector (ND) from Fermilab's Neutrinos at the Main Injector (NuMI) beam between November 2014 and February 2024, corresponding to an exposure of $25.5\times 10^{20}$ protons on target. The ND is a segmented tracking calorimeter located 800~m from the beam target, receiving neutrinos with a mean energy of 2~GeV. A multi-pass background reduction strategy is implemented, including the development of a novel dimuon-specific tracking technique. Trident candidates are identified using a boost ed decision tree classifier trained on simulated signal and background events. Limited background Monte Carlo statistics necessitate the use of functional fits to sideband data, which are extrapolated to estimate backgrounds in the signal region. The unblinded data contain 9 trident-like events, with an estimated background of 5.66 $\pm$ 5.15 events. This yields a best fit estimate of 3.34 tridents compared to the Standard Model prediction of 4.66. A profiled Feldman-Cousins method is used to determine a 90\% confidence interval of [0,9.1] on the number of signal events, corresponding to an upper limit of 1.95$\times$ the Standard Model prediction. This result represents the lowest energy search for trident events to date, and the first experimental contribution to the process in 27 years.

Bowles, Reed Scott [Indiana U.]↗

Off-the-shelf real-time monitoring of satellite constellations in a visual 3-D environment

The multimission spacecraft analysis system (MSAS) data monitor is a generic software product for future real-time data monitoring and analysis. The system represents the status of a satellite constellation through the shape, color, motion and position of graphical objects floating in a three dimensional virtual reality environment. It may be used for the monitoring of large volumes of data, for viewing results in configurable displays, and for providing high level and detailed views of a constellation of monitored satellites. It is considered that the data monitor is an improvement on conventional graphic and text-based displays as it increases the amount of data that the operator can absorb in a given period, and can be installed and configured without the requirement for software development by the end user. The functionality of the system is described, including: the navigation abilities; the representation of alarms in the cybergrid; limit violation; real-time trend analysis, and alarm status indication.

Schwuttke, Ursula M.↗

Neural network-based model of galaxy power spectrum: fast full-shape galaxy power spectrum analysis

ABSTRACT We present a neural network-based emulator for the galaxy redshift-space power spectrum that enables several orders of magnitude acceleration in the galaxy clustering parameter inference, while preserving 3$\sigma$ accuracy better than 0.5 per cent up to $k_{\mathrm{max}}$ = 0.25 $\, h\text{Mpc}^{-1}$ within Lambda-cold dark matter ($\Lambda$CDM) and around 0.5 per cent $w_0$–$w_a$CDM. Our surrogate model only emulates the galaxy bias-invariant terms of one-loop perturbation theory predictions, these terms are then combined analytically with galaxy bias terms, counter-terms, and stochastic terms in order to obtain the non-linear redshift-space galaxy power spectrum. This allows us to avoid any galaxy bias prescription in the training of the emulator, which makes it more flexible. Moreover, we include the redshift $z \in [0,1.4]$ in the training which further avoids the need for re-training the emulator. We showcase the performance of the emulator in recovering the cosmological parameters of $\Lambda$CDM by analysing the suite of 25 AbacusSummit simulations that mimic the Dark Energy Spectroscopic Instrument luminous red galaxies at $z=0.5$ and 0.8, together as the emission line galaxies at $z=0.8$. We obtain similar performance in all cases, demonstrating the reliability of the emulator for any galaxy sample at any redshift in $0 \lt z \lt 1.4$. We will make our emulator public at github repository.

Trusov, Svyatoslav (ORCID:0000000224146720)↗

NASA automatic subject analysis technique for extracting retrievable multi-terms (NASA TERM) system

Current methods for information processing and retrieval used at the NASA Scientific and Technical Information Facility are reviewed. A more cost effective computer aided indexing system is proposed which automatically generates print terms (phrases) from the natural text. Satisfactory print terms can be generated in a primarily automatic manner to produce a thesaurus (NASA TERMS) which extends all the mappings presently applied by indexers, specifies the worth of each posting term in the thesaurus, and indicates the areas of use of the thesaurus entry phrase. These print terms enable the computer to determine which of several terms in a hierarchy is desirable and to differentiate ambiguous terms. Steps in the NASA TERMS algorithm are discussed and the processing of surrogate entry phrases is demonstrated using four previously manually indexed STAR abstracts for comparison. The simulation shows phrase isolation, text phrase reduction, NASA terms selection, and RECON display.

Kirschbaum, J.↗

What Went Wrong: A Survey of Wildfire UAS Mishaps through Named Entity Recognition

Increasingly, unmanned aircraft systems (UAS) are being applied to wildfire incidents for tasks such as mapping, aerial ignition, and delivery. As a result, aviation incident reporting systems for wildfires are beginning to accumulate data related to UAS mishaps in wildfire response. In this research, we apply state-of-the-art natural language processing (NLP) techniques to develop a custom Named Entity Recognition (NER) model which extracts entities relevant to safety analysts. The custom NER model is built by fine-tuning an existing Bidirectional Encoder Representations from Transformers (BERT) model, resulting in a generalizable NER model that can extract engineering relevant entities including failure modes, causes, effects, control processes, and recommendations from failure-relevant text. This model performs passably, with a weighted average f1 score of 0.33 across entity types, indicating more labeled training data is needed. Extracted entities are used to form a Failure Modes and Effects Analysis (FMEA)-style survey of wildfire UAS mishaps reported using the SAFECOM system. Similar mishaps are manually clustered and reported as single rows within an FMEA. Foreach cluster, we compute frequency, severity, and overall riskin accordance with FAA standards. This methodology can beapplied as part of a broader safety management system totrack trends in mishaps (e.g., likelihood, severity) and discoverknowledge (e.g., causes, effects) that can be utilized to improvesafety outcomes and system performance.

Machine Learning↗

Leveraging BERT and Network-Based Attention Analysis for Identifying Treatment Milestones in EHRs

This study introduces a sophisticated data-driven framework for analyzing Electronic Health Records (EHRs) using transformer-based models to identify and disentangle overlapping treatment contexts. The framework leverages a preprocessing pipeline that transforms structured procedural codes into semantically enriched descriptive text, enabling the use of attention mechanisms to cluster medical events into treatment milestones—cohesive and distinct components of care processes. The methodology is rigorously validated using synthetic datasets derived from the MIMIC-III database, designed to simulate the heterogeneity and overlapping procedural contexts characteristic of real-world EHR scenarios. Quantitative evaluation highlights the framework’s robustness in disentangling concurrent care pathways, with attention metrics and unsupervised clustering approaches demonstrating the ability to preserve intra-context relationships while distinguishing inter-context dependencies. By addressing challenges inherent in data heterogeneity, this approach provides a foundation for uncovering complex treatment patterns, advancing clinical decision-making, and optimizing resource allocation in diverse healthcare environments.

Kim, Minsu [ORNL] (ORCID:0000000224185535)↗

STS propellant scavenging systems study. Part 2, volume 2: Cost and WBS/dictionary

Presented are the results of the cost analysis performed to update and refine the program phase C/D cost estimates for a Shuttle Derived Vehicle (SDV) tanker. The SDV tanker concept is an unmanned cargo vehicle incorporating a set of propellant tanks in the vehicle's payload module. The tanker will be used to meet the demand for a cryogenic propellant supply in orbit. The propellant tanks are delivered to a low Earth orbit or to an orbit in the vicinity of the Space Station. The intent of the economic analysis is to provide NASA with economic justification for the propellant scavenging concept that minimizes the total Space Transportation System life cycle cost. The detailed costs supporting the concept selection process are presented with descriptive text to aid in forecasting the phase C/D project and program planning. Included are all propellant scavenging costs as well as all SDV, STS and Orbital Maneuvering Vehicle charges to deliver the propellants to the Space Station.

Williams, Frank L.↗

Methods of applied dynamics

The monograph was prepared to give the practicing engineer a clear understanding of dynamics with special consideration given to the dynamic analysis of aerospace systems. It is conceived to be both a desk-top reference and a refresher for aerospace engineers in government and industry. It could also be used as a supplement to standard texts for in-house training courses on the subject. Beginning with the basic concepts of kinematics and dynamics, the discussion proceeds to treat the dynamics of a system of particles. Both classical and modern formulations of the Lagrange equations, including constraints, are discussed and applied to the dynamic modeling of aerospace structures using the modal synthesis technique.

Rheinfurth, M. H.↗

System engineering toolbox for design-oriented engineers

This system engineering toolbox is designed to provide tools and methodologies to the design-oriented systems engineer. A tool is defined as a set of procedures to accomplish a specific function. A methodology is defined as a collection of tools, rules, and postulates to accomplish a purpose. For each concept addressed in the toolbox, the following information is provided: (1) description, (2) application, (3) procedures, (4) examples, if practical, (5) advantages, (6) limitations, and (7) bibliography and/or references. The scope of the document includes concept development tools, system safety and reliability tools, design-related analytical tools, graphical data interpretation tools, a brief description of common statistical tools and methodologies, so-called total quality management tools, and trend analysis tools. Both relationship to project phase and primary functional usage of the tools are also delineated. The toolbox also includes a case study for illustrative purposes. Fifty-five tools are delineated in the text.

Goldberg, B. E.↗

A Flight Rule Checker for the LADEE Lunar Spacecraft

As part of the design of a space mission, an important part is the design of so-called flight rules. Flight rules express constraints on various parts and processes of the mission, that if followed, will reduce the risk of failure. One such set of flight rules constrain the format of command sequences regularly (e.g. daily) sent to the spacecraft to con- trol its next near term behavior. We present a high-level view of the automated flight rule checker Frc for checking command sequences sent to NASA’s LADEE Lunar mission spacecraft, used throughout its entire mission. A command sequence is in this case essentially a program (a sequence of commands) with no loops or conditionals, and it can there- fore be verified with a trace analysis tool. Frc is implemented using the TraceContract runtime verification tool, an internal Scala DSL for checking event sequences against “formal specifications”. The paper illustrates this untraditional use of runtime verification in a real con- text, with strong demands on the expressiveness and flexibility of the specification language, illustrating the advantages of an internal DSL.

Kurklu, Elif↗

Tool for Generation of MAC/GMC Representative Unit Cell for CMC/PMC Analysis

This document describes a recently developed analysis tool that enhances the resident capabilities of the Micromechanics Analysis Code with the Generalized Method of Cells (MAC/GMC) 4.0. This tool is especially useful in analyzing ceramic matrix composites (CMCs), where higher fidelity with improved accuracy of local response is needed. The tool, however, can be used for analyzing polymer matrix composites (PMCs) as well. MAC/GMC 4.0 is a composite material and laminate analysis software developed at NASA Glenn Research Center. The software package has been built around the concept of the generalized method of cells (GMC). The computer code is developed with a user friendly framework, along with a library of local inelastic, damage, and failure models. Further, application of simulated thermomechanical loading, generation of output results, and selection of architectures to represent the composite material have been automated to increase the user friendliness, as well as to make it more robust in terms of input preparation and code execution. Finally, classical lamination theory has been implemented within the software, wherein GMC is used to model the composite material response of each ply. Thus, the full range of GMC composite material capabilities is available for analysis of arbitrary laminate configurations as well. The primary focus of the current effort is to provide a graphical user interface (GUI) capability that generates a number of different user-defined repeating unit cells (RUCs). In addition, the code has provisions for generation of a MAC/GMC-compatible input text file that can be merged with any MAC/GMC input file tailored to analyze composite materials. Although the primary intention was to address the three different constituents and phases that are usually present in CMCs-namely, fibers, matrix, and interphase-it can be easily modified to address two-phase polymer matrix composite (PMC) materials where an interphase is absent. Currently, the tool capability includes generation of RUCs for square packing, hexagonal packing, and random fiber packing as well as RUCs based on actual composite micrographs. All these options have the fibers modeled as having a circular cross-sectional area. In addition, a simplified version of RUC is provided where the fibers are treated as having a square cross section and are distributed randomly. This RUC facilitates a speedy analysis using the higher fidelity version of GMC known as HFGMC. The first four mentioned options above support uniform subcell discretization. The last one has variable subcell sizes due to the primary intention of keeping the RUC size to a minimum to gain the speed ups using the higher fidelity version of MAC. The code is implemented within the MATLAB (The Mathworks, Inc., Natick, MA) developmental framework; however, a standalone application that does not need a priori MATLAB installation is also created with the aid of the MATLAB compiler.

Materials Engineering↗