Search NASA⌕ Search

SEARCH · Search NASA

Results for “count data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Multiple Changepoint Detection for Non‐Gaussian Time Series

ABSTRACT This article combines methods from existing techniques to identify multiple changepoints in non‐Gaussian autocorrelated time series. A transformation is used to convert a Gaussian series into a non‐Gaussian series, enabling penalized likelihood methods to handle non‐Gaussian scenarios. When the marginal distribution of the data is continuous, the methods essentially reduce to the change of variables formula for probability densities. When the marginal distribution is count‐oriented, Hermite expansions and particle filtering techniques are used to quantify the scenario. Simulations demonstrating the efficacy of the methods are given and two data sets are analyzed: 1) the proportion of home runs hit by Major League Baseball batters from 1920 to 2023 and 2) a six‐dimensional series of tropical cyclone counts from the Earth's basins of generation from 1980 to 2023. In the first series, beta marginal distributions are used to describe the proportions; in the second, Poisson marginal distributions seem appropriate.

Lund, Robert [Department of Statistics University ↗

Data for: Climatic Imprint on Interfacially-Controlled Platinum-Palladium Resources

Data package for manuscript "Climatic Imprint on Interfacially-Controlled Platinum-Palladium Resources" by Emily G. Wright, Ivey Wang, Yihang Fang, Elaine D. Flynn, and Jeffrey G. Catalano. This dataset contains adsorption results from experiments designed to investigate the effect of chloride on Pd(II) adsorption to goethite and Pt(II) adsorption to hematite and goethite, including lab experiments, X-ray absorption fine structure spectroscopy, and models of retention within a laterite. See the associated manuscript for full methods information. The file "Wright2025_PtAds_data.csv" contains the target starting Pt concentration (uM), final aqueous Pt and associated error (in uM), calculated adsorbed Pt and associated error (in umol/m2), target and measured aqueous chloride (mM), target aqueous nitrate (mM), final pH, and mineral concentration/loading (g/L). Associated mineral-free controls (mineral loading = 0 g/L) are included; the aqueous Pd error was not calculated and chloride was not measured in every sample. These data appear in Figures 1, S3, S4, S5, S20, and S22 in the associated manuscript. The file "Wright2025_PdAds_data.csv" contains the target starting Pd concentration (uM), final aqueous Pd and associated error (in uM), calculated adsorbed Pd and associated error (in umol/m2), target and measured aqueous chloride (mM), and mineral concentration/loading (g/L). Associated mineral-free controls (mineral loading = 0 g/L) are included; the aqueous Pd error was not calculated and chloride was not measured in every sample. These data appear in Figures 1, S3, S4, S5, and S20 in the associated manuscript. The file "Wright2025_MineralBatches_data.csv" contains the mineral identity and BET specific surface area (m2/g) for every mineral batch synthesized and used in experiments. The annealing time used is listed for hydrothermally annealed goethite. These data appear in Table S2 in the associated manuscript. The file "Wright2025_XRD_data.csv" contains the XRD patterns for every mineral batch synthesized as the counts as a function of two theta (in degrees). See "Wright2025_MineralBatches_data.csv" for more details on specific mineral batches. These data appear in Figure S2 in the associated manuscript. The file "Wright 2025_ZetaPotential_data.csv" contains the measured zeta potentials for samples of goethite (batch G2) at pH 4 the presence of varying amounts of sodium chloride. These data appear in Table S3 in the associated manuscript. The file "Wright2025_XAFSSamples_data.csv" contains the specific mineral batch, measured final aqueous Pd or Pt (uM), measured final aqueous chloride (mM), and estimated adsorbed Pd or Pt (umol/m2) of all XAFS samples. These data appear in Tables S4, S7, S8, and S10 in the associated manuscript. The files "Wright2025_PdXAFS_data.csv" and "Wright2025_PtXAFS_data.csv" contain the normalized spectra of Pd and Pt, respectively, adsorbed to minerals at varying chloride concentrations. See "Wright2025_XAFSSamples_data.csv" for a guide to sample names. Note that "05" in a sample name is equivalent to "0.5". These data appear in Figures 2, S6, S7, S8, S12, S13, and S14 in the associated manuscript. The file "Wright2025_LateriteProfileProfileModelParameters_data.csv" include the ratio of hematite to hematite and goethite in two synthetic, modeled profiles, as well as the modeled surface areas of goethite and hematite as a function of relative depth within the modeled weathering zone. These data were used, in conjunction with equations presented in the paper, to calculate the theoretical concentrations of Pd and Pt (and the resulting Pt/Pd ratio) within the profiles. These data appear in Figure 3 in the associated manuscript. The file "Wright2025_Imagery_data.zip" is a zipped folder containing the TEM and STEM images appear in Figures S18 and S19. Individual files are labeled as either STEM (Fig. S18) or TEM (Fig. S19) with a letter representing the part of the multipart figure.

58 GEOSCIENCES↗

Statistical analysis of the performance and long-term stability of unquenched LSC standards ( 3 H, 14 C) used for radiochemical measurements

Multiple sets of toluene-based unquenched standards, procured over the last 3 decades, were analyzed to determine their stability over time. A statistical analysis was performed to provide insight into the variability for tSIE and counting efficiency measurements of 3 H and 14 C isotopes. Our data suggests that standards remain viable even 30 years after the manufacturer’s recommended expiration date. Overall, these standards have shown no statistically significant sign of performance degradation. We have proven through statistical analysis that these standards can provide comparable performance over time well past their manufactured expiration date.

LSC standards↗

High-Multiplicity Muon Airshower Analysis at NOvA Far Detector

We process and analyze muon airshower data from the NOvA far detector using various image processing algorithms, such as Fast Fourier Transformation, and Hough line transformation. From the processed event images, we calculate multiple parameters for our study. We are looking for physics features, including East-West Asymmetry, anisotropies in right ascension, and seasonal variation. Additionally, we have developed an algorithm to count the multiplicity of muons in the airshower events using the single muon data.

Lima, Aklima Khanam [Syracuse U.]↗

Archival records housed at USTUR support radium dial worker dosimetry

The American radium dial worker (RDW) cohort of over 3200 persons is being revisited as part of the Million Person Study (MPS) to include a modern approach to RDW dosimetry. An exceptional source of data and contextualization in this project is an extensive collection of electronic records (digitized from existing microfilm and microfiche) housed at the United States Transuranium and Uranium Registries (USTUR). Although the type, extent, and quality (e.g. legibility) of record(s) varies between individuals, the remarkable occupational, medical and demographic data include in vivo radiation measurements (e.g. radon breath, whole body counts), autopsy results, medical records (including copies of radiographs), interviews over the years, and correspondence. Of particular dosimetric interest are the details of radiation measurements. For example, there are some instances where hand-written and transcribed values are both available, along with notes providing context for why a particular measurement in a series of measurements was chosen to assign an intake, or if there were concerns about a particular measurement. Born prior to 1935, RDW have nearly all passed away. Thus, the updated dosimetry, especially for the skeletal tissues, will allow the correlation of lifetime cumulative dose with radiation risk. Here we review typical information available in this collection of historical records and highlight some interesting finds. Additionally, we discuss the relevance to current and ongoing work related to updating the dosimetry of the RDW in the MPS, including providing an example of the usefulness of information contained in these records. The RDW cohort provides a unique historical perspective on occupational exposure to radium, making it a valuable dataset for understanding long-term health effects and improving current radiation protection standards.

Million Person Study↗

On the Prospect of Chemically Transferable Coarse-Grained Electronic Models for Soft Materials

Electronic coarse-graining (ECG) methods predict quantum-mechanical electronic properties directly from coarse-grained (CG) molecular configurations, enabling electronic predictions at mesoscale length scales. Here, we present a diagnostic assessment of the feasibility of chemically transferable ECG models across a broad polymer-relevant chemical space using all-atom, united-atom, and Martini-scale representations. While high-resolution ECG models achieve near-quantitative accuracy, we show that chemically transferable ECG at the Martini resolution fails because the CG force field does not sample the same configurational distribution of local molecular structure as that underlying the DFT-parameterized ECG model. We demonstrate that our proposed Element-Count-Label (ECL) representation, which augments Martini beads with explicit stoichiometric data, significantly improves chemical generalization across diverse polymer chemistries. However, we find that even with improved chemical resolution, the model cannot recover electronic property distributions that are absent from the configurational space sampled by the CG force field. These results demonstrate that chemically transferable ECG requires future Martini-like force fields to explicitly preserve quantum chemistry–compatible local molecular structure in addition to thermodynamic and structural fidelity.

Kidder, Katherine M [Department of Chemistry; Univ↗

Soil biogeochemical properties and metrics of tree-mycorrhizal dominance for a 25-Ha forest in South Central Indiana, USA.

This data package contains a dataset used in the papers “Seeing the forest for all the trees: Mycorrhizal-associated nutrient economies are modulated by stem density and the synchrony between overstory and understory communities” and “Mycorrhizal associations of tree species influence soil nitrogen dynamics via effects on soil acid–base chemistry”. Four csv files are included along with a dataset. The dataset features chemical soil properties for a single sampling campaign within the 25 Ha Lilly-Dickey Woods Smithsonian Forest Global Earth Observatory (ForestGEO) plot in South Central Indiana, USA (ldw_dat_raw.csv). Also included are separate files focused on pH (pH_data.csv), carbon and nitrogen (CN_data.csv), and nitrification rates (Nitrification_data.csv). These variables are commonly associated with the tree-mycorrhizal dominance of forest stands. In these data subsets, each soil variable was matched to a 10 meter radius neighborhood wherein metrics of tree-mycorrhizal dominance (basal area, stem count, importance value, etc.) were calculated. Models between these soil variables and dominance metrics were used to investigate how different assessments of mycorrhizal associated nutrient economies (MANE) capture these relationships. This research was performed as a part of the Smithsonian ForestGEO project. This data package can be used to explore spatial variability in soil chemistry within a mature hardwood forest, or it can be combined with the included tree data, other fine-scale spatial information, or other tree inventory data for the site to evaluate how soil chemistry varies with tree community composition or edaphic or topographic properties.

Craig, Matthew [ORNL] (ORCID:0000000288907920)↗

Measurement of charged hadron multiplicity in Au + Au collisions at $\sqrt{s_{NN}}$ = 200 GeV with the sPHENIX detector

The pseudorapidity distribution of charged hadrons produced in Au + Au collisions at a center-of-mass energy of $\sqrt{s_{NN}}$ = 200 GeV is measured using data collected by the sPHENIX detector. Charged hadron yields are extracted by counting cluster pairs in the inner and outer layers of the Intermediate Silicon Tracker, with corrections applied for detector acceptance, reconstruction efficiency, combinatorial pairs, and contributions from secondary decays. The measured distributions cover |η| < 1.1 across various centralities, and the average pseudorapidity density of charged hadrons at mid-rapidity is compared to predictions from Monte Carlo heavy-ion event generators. This result, featuring full azimuthal coverage at mid-rapidity, is consistent with previous experimental measurements at the Relativistic Heavy Ion Collider, thereby supporting the broader sPHENIX physics program.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Leveraging Large Language Models for Real-World Data Evidence: A Framework for Automated Treatment Extraction and Data Harmonization

Background: The ability to comprehensively collect treatment information from cancer patient medical records would enable studies to evaluate real-world benefits and risks tied to specific treatments. Currently, it is difficult to system- atically collect high-quality treatment information because it is often stored in unstructured text. Manually extracting and standardizing drug and regimen data is time-intensive. Recent advances in large language models (LLMs) offer a potential solution for automated extraction of structured treatment information from clinical text. Objective: This study systematically evaluates the utility of four LLMs from the Llama family for automated extraction of oncology treatment information from clinical text. This information can guide researchers using cancer registry data to provide insights into cancer care and outcomes beyond clinical trials. Methods: Four instruction-tuned Llama models with varying parameter counts (1B, 3B, 8B, and 70B) were evaluated for their ability to extract treatment information from clinical documents. A unified oncology knowledge base integrating seven major public data sources was developed to standardize and normalize extracted entities—a critical step for harmonizing data from diverse sources. Extracted treatment data were compared against expert-annotated ground truth. Model performance was assessed using accuracy metrics (Precision, Recall, F1-Score) and opera- tional feasibility metrics, including processing speed and structural compliance of the output. Results: A strong positive correlation was observed between model size and extraction accuracy. F1-score improved from 0.609 for the 1B model to 0.710 (3B), 0.807 (8B), and 0.828 (70B). While larger models demonstrated superior accuracy and compliance, they incurred higher computational costs. The modest performance difference between 8B and 70B suggests diminishing returns with increasing model size. Conclusions: LLMs represent a viable technology for automating oncology treatment extraction. The 8B-parameter model emerged as a highly effective option, balancing high accuracy and computational efficiency. Selecting an appropriate LLM for deployment in cancer registries involves a trade-off between desired accuracy and available operational resources. Harmonizing extracted entities with the oncology knowledge base facilitates standardized integration into common data models, enhancing data quality for real-world evidence analyses.

artificial intelligence↗

Uranium particle age dating, aggregation, and model age best estimators

We present important aspects of uranium particle age dating by Large-Geometry Secondary Ion Mass Spectrometry (LG-SIMS) that can introduce bias and increase model age uncertainties, especially for small, young, and/or low-enriched particles. This metrology is important for applications related to International Nuclear Safeguards. We explore influential factors related to model age estimation, including the effects of evolving surface chemistry on inter-element measurements of particles (e.g., Th and U), detector background, and aggregation methods using simulated and actual particle samples. We introduce a new model age estimator, called “mid68”, that supplements 95% confidence intervals, providing a “best estimate” and uncertainty about the most likely age. The mid68 estimator can be calculated using the Feldman and Cousins method or Bayesian methods and provides a value with a symmetric uncertainty that can be used for calculations and approximate aggregation of processed model age values when the raw data and correction factors are not available. For particles yielding low 230 Th counts amidst nonzero detector background, their underlying model age probability distributions are asymmetric, so the mid68 estimator provides additional robust information regarding the underlying model age likelihood. This study provides a comprehensive and timely examination of critical aspects of uranium particle age dating as more laboratories establish particle chronometry capabilities.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

The SRG/eROSITA All-Sky Survey: Dark Energy Survey year 3 weak gravitational lensing by eRASS1 selected galaxy clusters

Context. Number counts of galaxy clusters across redshift are a powerful cosmological probe if a precise and accurate reconstruction of the underlying mass distribution is performed – a challenge called mass calibration. With the advent of wide and deep photometric surveys, weak gravitational lensing (WL) by clusters has become the method of choice for this measurement. Aims. We measured and validated the WL signature in the shape of galaxies observed in the first three years of the Dark Energy Survey (DES Y3) caused by galaxy clusters and groups selected in the first all-sky survey performed by SRG (Spectrum Roentgen Gamma)/eROSITA (eRASS1). These data were then used to determine the scaling between the X-ray photon count rate of the clusters and their halo mass and redshift. Methods. We empirically determined the degree of cluster member contamination in our background source sample. The individual cluster shear profiles were then analyzed with a Bayesian population model that self-consistently accounts for the lens sample selection and contamination and includes marginalization over a host of instrumental and astrophysical systematics. To quantify the accuracy of the mass extraction of that model, we performed mass measurements on mock cluster catalogs with realistic synthetic shear profiles. This allowed us to establish that hydrodynamical modeling uncertainties at low lens redshifts (z < 0.6) are the dominant systematic limitation. At high lens redshift, the uncertainties of the sources’ photometric redshift calibration dominate. Results. With regard to the X-ray count rate to halo mass relation, we determined its amplitude, its mass trend, the redshift evolution of the mass trend, the deviation from self-similar redshift evolution, and the intrinsic scatter around this relation. Conclusions. The mass calibration analysis performed here sets the stage for a joint analysis with the number counts of eRASS1 clusters to constrain a host of cosmological parameters. We demonstrate that WL mass calibration of galaxy clusters can be performed successfully with source galaxies whose calibration was performed primarily for cosmic shear experiments, opening the way for the cluster cosmological exploitation of future optical and NIR surveys like Euclid and LSST.

79 ASTRONOMY AND ASTROPHYSICS↗

Evaluating the Nation's Pipeline Infrastructure with NETL's Advanced Infrastructure Integrity Model (AIIM)

This poster is a part of BIL-EDX4CCS Task 36: Advanced Infrastructure Integrity Modeling to Evaluate Existing Energy Infrastructure Reusability and Risk, the goal of which is to produce a smart tool that will assess existing energy infrastructure reusability and risk using the Advanced Infrastructure Integrity Model (AIIM). This model forecasts lifespan and potential risk using a multitude of factors such as incidents reports, structural characteristics, and the surrounding environment. The project aims to provide scientific insights for a better understanding of carbon storage (CS), potential to support CS stakeholder needs, national decarbonization, and mitigating climate change. AIIM will utilize an energy infrastructure database as its input, developed by acquiring publicly available data as well as NETL derived products. These resources include incidents, geohazards, and infrastructure variables. Soil data in the form of rasters and pipeline incident reports were processed and a script was developed to count the number of times features such as roads, railroads, and rivers intersected with pipeline segments which were then converted to points. Distance to oil and natural gas wells, petroleum ports, intermodal freight facilities, and geologic structures were also calculated. After data preparation and quality control was completed, the data was integrated into the pipeline points. Once models are complete, a smart tool will be created in the form of an online dashboard.

Malay, Caleb↗

PhotonIDs: ML-Powered Photon Identification System for Dark Count Elimination

Reliable single photon detection is the foundation for practical quantum communication and networking. However, today's superconducting nanowire single photon detector(SNSPD) inherently fails to distinguish between genuine photon events and dark counts, leading to degraded fidelity in long-distance quantum communication. In this work, we introduce PhotonIDs, a machine learning-powered photon identification system that is the first end-to-end solution for real-time discrimination between photons and dark count based on full SNSPD readout signal waveform analysis. PhotonIDs ~demonstrates: 1) an FPGA-based high-speed data acquisition platform that selectively captures the full waveform of signal only while filtering out the background data in real time; 2) an efficient signal preprocessing pipeline, and a novel pseudo-position metric that is derived from the physical temporal-spatial features of each detected event; 3) a hybrid machine learning model with near 98% accuracy achieved on photon/dark count classification. Additionally, proposed PhotonIDs ~ is evaluated on the dark count elimination performance with two real-world case studies: (1) 20 km quantum link, and (2) Erbium ion-based photon emission system. Our result demonstrates that PhotonIDs ~could improve more than 31.2 times of signal-noise-ratio~(SNR) on dark count elimination. PhotonIDs ~ marks a step forward in noise-resilient quantum communication infrastructure.

Linne, Karl C. [Chicago U.] (ORCID:000900091870358↗

High-count-rate effects in event processing for the XRISM/Resolve X-ray microcalorimeter. II. Energy scale and resolution in orbit

The Resolve instrument on the X-ray Imaging and Spectroscopy Mission (XRISM) uses a 36 pixel microcalorimeter designed to deliver high-resolution, non-dispersive X-ray spectroscopy. Although it is optimized for extended sources with low count rates, Resolve observations of bright point sources are still able to provide unique insights into the physics of these objects, as long as high-count-rate effects are addressed in the analysis. These effects include the loss of exposure time for each pixel, changes in the energy scale, and changes in the energy resolution. To investigate these effects under realistic observational conditions, we observed the bright X-ray source, the Crab Nebula, with XRISM at several offset positions with respect to the Resolve field of view and with continuous illumination from 55 Fe sources on the filter wheel. For the spectral analysis, we excluded data where exposure-time loss was too significant to ensure reliable spectral statistics. The energy scale at 6 keV shows a slight negative shift in the high-count-rate regime. The energy resolution at 6 keV worsens as the count rate in electrically neighboring pixels increases, but can be restored by applying a nearest-neighbor coincidence cut (“cross-talk cut”). We examined how these effects influence the observation of bright point sources, using GX 13+1 as a test case, and identified an eV-scale energy offset at 6 keV between the inner (brighter) and outer (fainter) pixels. Users who seek to analyze velocity structures on the order of tens of km s–1 should account for such high-count-rate effects. These findings will aid in the interpretation of Resolve data from bright sources and provide valuable considerations for designing and planning for future microcalorimeter missions.

X-rays: general↗

1990 The Twin Cities Metropolitan Area Travel Behavior Inventory

The Twin Cities Metropolitan Area 1990 Travel Behavior Inventory, or home interview survey, was conducted by Twin Cities Metropolitan Council—Saint Paul to document how Twin Cities residents use the streets, highways, and transit services in the region. The survey collected demographic, socioeconomic, and travel data for 9,476 households. Respondents were asked to record their travel and activities for a 24-hour period, in which 98,535 trips were documented. In addition to the home interview survey, the study included an establishment survey, a transit survey, external station traffic counts, and an external station origin/destination survey.

1Hz data↗

Multi-resonant switched capacitor power converter architecture

A switched-capacitor (SC) network in an SC converter is controlled to operate at varying resonant modes to achieve high conversion ratio efficiency, at a low circuit component count. These power converters are suited to numerous application areas including improving energy efficiency of data centers. A family of resonant switched capacitor (SC) converters with multiple operating phases are presented “Multi-Resonant SC Converters”. Described in detail are an 8-to-1 Multi-Resonant-Doubler (MRD) converter and a 6-to-1 Cascaded Series-Parallel (CaSP). The topology of these converters make them amenable to combining like units in parallel toward reaching higher power levels.

Ye, Zichao↗

All loop scattering as a counting problem

Abstract This is the first in a series of papers presenting a new understanding of scattering amplitudes based on fundamentally combinatorial ideas in the kinematic space of the scattering data. We study the simplest theory of colored scalar particles with cubic interactions, at all loop orders and to all orders in the topological ’t Hooft expansion. We find novel integral formulas for the amplitudes of this theory, with no trace of the conventional sum over Feynman diagrams, but instead determined by a beautifully simple counting problem attached to any order of the topological expansion. These results represent a significant step forward in the decade-long quest to formulate the fundamental physics of the real world in a radically new language, where the rules of spacetime and quantum mechanics, as reflected in the principles of locality and unitarity, are seen to emerge from deeper mathematical structures.

1/N Expansion↗

SPRUCE Root Tip and Ectomycorrhizal Fungi Colonization Measurements from Ingrowth Cores, 2017

This data set contains root tip and ectomycorrhizal fungi colonization measurements taken from ingrowth cores from the SPRUCE experiment (Hanson et al. 2017) that were deployed during the 2017 growing season (2017-06 to 2017-10-01). This study explored the relationship between warming treatments and fine-root growth. Increased fine-root growth may increase root exudates and accelerate turnover, representing an underlying mechanism for peat decomposition through priming, as exudates provide a labile carbon source to the microbial community. Roots of two tree species were studied: an evergreen conifer Picea mariana (black spruce) and a deciduous conifer Larix laricina (tamarack). Measurements include root tips counts and densities by tree species and the abundance of ectomycorrhizal colonization on root tips. This dataset contains one data file in comma separate (.csv) format. Additional metadata are provided: one data dictionary and a file-level metadata file in comma separate (.csv) format and a user guide in PDF (*.pdf) format.

black spruce [Picea mariana]↗