Search NASA⌕ Search

SEARCH · Search NASA

Results for “data return”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Multi-modal dynamic radiography using short-pulse laser-generated probe beams

Radiography is an important tool for the interrogation of dynamic experiments in the fields of dynamic properties of materials, and in condensed matter, high explosive, and high-energy-density physics. Multi-modal radiography advances the hypothesis that combining the information delivered by multiple radiographic modalities can lead to more constrained (improved) “reconstruction” of the scene than can be obtained from a single probe. We identify four modalities: multi-probe, time sequence, multi-view, and multi-messenger. Multi-probe radiography is a promising candidate for a next-generation dynamic radiographic facility. High-energy X-rays are the most frequently used probe for dynamic radiography, although recent developments show the utility of proton (pRad), electron (eRad), and neutron probe beams. Because each probing species interacts with material in the radiographic scene through quantitatively different mechanisms, each returns independent information about the scene, which can add extra constraints to the reconstruction process. How to conduct detailed, quantitative “co-analysis” of multiple data streams remains an area of active research. Multi-beam, short-pulse, laser-generated probes offer sufficient dose, an appropriate spectrum, and appropriate spatio-temporal resolution to produce high-quality dynamic radiographs. This paper reports on technology development to advance the state of the art of multi-modal/multi-probe radiography and the pursuit of both deterministic and inferential (AI/ML assisted) co-analysis methodologies to produce more constrained reconstructions from multi-modal data.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Automatic Loss Factor Modeling and Attribution on Unlabeled PV Energy Data

We present a novel approach for modeling the loss factors of photovoltaic power generation systems (PV systems). This method is a white-box machine learning model built on convex optimization that is fast, interpretable, and auditable. It takes as an input the measured daily energy produced by the system, over a multi-year period, and returns a multiplicative decomposition model of the daily energy signal and full attribution of the total energy loss to each feature. The methods section of this paper has two major components: (1) the description of the signal decomposition (SD) model, expressed in the SD framework, and (2) the attribution of total energy losses via Shapley values. We validate the method on synthetic and open-source data sets and compare to similar methods from the literature.

artificial intelligence↗

Reconciling the Different Apparent Decay Pathways for Iron(III) Tetraphenylporphyrin Chloride: Evidence for Excited State Branching

Detailed understanding of the photophysical properties of metalloporphyrins is key to rationally exploiting them in a variety of applications ranging from photocatalysis to opto-magnetics. Previous studies of the ferric-tetraphenylporphyrin chloride have provided contradictory descriptions of the excited state evolution. Optical transient-absorption suggested initial formation of a ππ* excited state, followed by ligand-to-metal charge-transfer (LMCT) from the porphyrin ring and then decay on the ∼2 ps time scale to a metal-centered excited state that had a lifetime of ∼15 ps. In contrast, femtosecond extreme ultraviolet transient-absorption at the Fe M 2,3 -edge, while agreeing on the initial formation of an LMCT state, found evidence that this decayed in ∼2 ps to the ground state. Here, we have used K-edge transient X-ray absorption and X-ray emission, together with time-resolved X-ray solution scattering to explore this system. Based on these data, we propose a new model, consistent with both the earlier and the current data, in which photoexcited FeTPPCl evolves through three different states on the LMCT manifold, with the ∼2 ps decay now seen to involve a branching between return to the ground state (∼70%) and formation of a long-lived LMCT state (∼30%).

Excited states↗

DMTN-326: Bulk Cutout Service Implementation Options

The bulk cutout service is a deliverable required to support Data Preview 2. This document explains our implementation options regarding how to perform the cutouts at scale and in what form the resulting cutouts should be returned.

79 ASTRONOMY AND ASTROPHYSICS↗

Metal Scrap Upcycling with Shear Assisted Processing and Extrusion (ShAPE)

The overarching objective of this project is to convert metal scraps, such as aluminum, titanium, and other alloys provided by the industry, into extruded tubing, wires, and rods. Upcycling of scrap will be accomplished using Shear Assisted Processing and Extrusion (ShAPE). This approach is a new solution for recycling. The specific aims of this project are as follows: 1. Receive metal scrap under a Material Transfer Agreement (MTA), in the form of billets, from select industry partners that meet the following requirements: outer diameter of 1.245 inches (+/-0.003 inches), inner diameter drilled with a 0.404-inch drill bit, and a length of 4.0 inches (+/-0.01 inches). The billet must be cast or compacted to greater than 98% density. If the industry partner does not have the capabilities to prepare the billets, PNNL can make introductions to third-party entities as needed. Industry partners will also provide a composition analysis as weight percent. 2. Extrude metal scrap via the ShAPE process at PNNL. 3. Evaluate and benchmark the extrudate material properties per the ASTM B557-15 – Testing Tubulars Standard or similar wire and rod standards. 4. Characterize the extrudate microstructure for any of the following: grain size, second phase composition, and texture. Additional testing may include corrosion per the ASTM B117 standard and electro-potential. 5. Deliver specimens to industry partners under an MTA for additional third-party evaluations. Industry partners will, in return, provide a non-proprietary report on any testing, including testing per the ASTM B557-15, ASTM B117, and other industry standards. 6. PNNL will develop data and insights to support intellectual property capture, a published non-proprietary technical report, research collaborations, and commercialization opportunities.

36 MATERIALS SCIENCE↗

Missing microbial eukaryotes and misleading meta-omic conclusions

Meta-omics is commonly used for large-scale analyses of microbial eukaryotes, including species or taxonomic group distribution mapping, gene catalog construction, and inference on the functional roles and activities of microbial eukaryotes in situ. Here, we explore the potential pitfalls of common approaches to taxonomic annotation of protistan meta-omic datasets. We re-analyze three environmental datasets at three levels of taxonomic hierarchy in order to illustrate the crucial importance of database completeness and curation in enabling accurate environmental interpretation. We show that taxonomic membership of sequence clusters estimates community composition more accurately than returning exact sequence labels, and overlap between clusters can address database shortcomings. Clustering approaches can be applied to diverse environments while continuing to exploit the wealth of annotation data collated in databases, and selecting and evaluating these databases is a critical part of correctly annotating protistan taxonomy in environmental datasets. We argue that ongoing curation of genetic resources is crucial in accurately annotating protists in in situ meta-omic datasets. Moreover, we propose that precise taxonomic annotation of meta-omic data is a clustering problem rather than a feasible alignment problem.

59 BASIC BIOLOGICAL SCIENCES↗

An absolute recoil-carbon polarimeter for polarized light-ions at the BNL Booster

We describe a recoil-carbon polarimeter for the BNL Booster synchrotron capable of measuring the transverse polarization of both the polarized proton beam and the polarized 3 He ion beam throughout the acceleration cycle. The instrument addresses a critical gap in the BNL polarized beam program: no independent polarization measurement currently exists in the injector chain between the 200 MeV Linac polarimeter and the AGS. The proposed new polarimeter will provide continuous, absolute polarization measurements at the Booster stage, supplying additional anchor points for the polarization transmission through the accelerator chain. The polarimeter 9 is based on elastic pC and 3 He C scattering from a thin internal carbon target, with detection of the recoil 12 C nucleus in a fixed six-station silicon detector ring. A single detector geometry provides continuous kinematic coverage from injection to extraction for both beam species without mechanical adjustment. For pC scattering, the available data and the Bonin parametrization provide a well-established basis for estimating the analyzing power and figure of merit over the Booster energy range; for 3 He C the sole experimental anchor is a measurement at 443 MeV, and the Booster polarimeter itself is identified as the instrument to map the analyzing power across the remainder of the acceleration ramp by a ramp-and-return calibration that anchors the absolute 3 He polarization scale to a few percent. Provision for future deuteron polarimetry is incorporated in the chamber design without modification to the existing geometry. With appropriate detector upgrades, the same recoil-carbon technique can be extended to deuteron beams, for which analyzing power data are already available, and, with future analyzing-power measurements, also to 6 Li and 7 Li beams, making the polarimeter a versatile instrument for the full range of light polarized ion species anticipated at BNL.

43 PARTICLE ACCELERATORS↗

Increased Occurrence of Large–Scale Windthrows Across the Amazon Basin

Convective storms with strong downdrafts create windthrows: snapped and uprooted trees that locally alter the structure, composition, and carbon balance of forests. Comparing Landsat imagery from subsequent years, we documented temporal and spatial variation in the occurrence of large (≥30 ha) windthrows across the Amazon basin from 1985 to 2020. Over 33 individual years, we detected 3179 large windthrows. Windthrow density was greatest in the central and western Amazon regions, with ~33% of all events occurring in ~3% of the monitored area. Return intervals for large windthrows in the same location of these “hotspot” regions are centuries to millennia, while over the rest of the Amazon they are >10,000 years. Our data demonstrate a nearly 4–fold increase in windthrow number and affected area between 1985 (78 windthrows and 6,900 ha) and 2020 (264 events and 32,170 ha), with more events of >500 ha size since 1990. Such extremely large events (>500 ha up to 2,543 ha) are responsible for interannual variation in the overall median (84 ± 5.2 ha; ±95% CI) and mean (147 ± 13 ha) windthrow area, but we did not find significant temporal trends in the size distribution of windthrows with time. Our results document increased damage from convective storms over the past 40 years in the Amazon, filling a gap in temporal records for tropical regions. Our publicly accessible large windthrow database provides a valuable tool for exploring dynamic conditions leading to damaging storms and their ecological impact on Amazon forests.

54 ENVIRONMENTAL SCIENCES↗

Harmonic analysis of discrete tracers of large-scale structure

It is commonplace in cosmology to analyze fields projected onto the celestial sphere, and in particular density fields that are defined by a set of points e.g. galaxies. When performing an harmonic-space analysis of such data (e.g. an angular power spectrum) using a pixelized map one has to deal with aliasing of small-scale power and pixel window functions. We compare and contrast the approaches to this problem taken in the cosmic microwave background and large-scale structure communities, and advocate for a direct approach that avoids pixelization. We describe a method for performing a pseudo-spectrum analysis of a galaxy data set and show that it can be implemented efficiently using well-known algorithms for special functions that are suited to acceleration by graphics processing units (GPUs). The method returns the same spectra as the more traditional map-based approach if in the latter the number of pixels is taken to be sufficiently large and the mask is well sampled. The method is readily generalizable to cross-spectra and higher-order functions. It also provides a convenient route for distributing the information in a galaxy catalog directly in harmonic space, as a complement to releasing the configuration-space positions and weights, and a route to spectral apodization. Finally, we make public a code enabling the application of our method to existing and upcoming datasets.

79 ASTRONOMY AND ASTROPHYSICS↗

ML-based Micro-CT SOFC Microstructure Models (from Kent 2026 Microstructural Augmentation paper)

Overview -------------------------- This repository contains datasets from the manuscript **"Enhanced Generalizability to Deep-Learning Quantification of 3D Microstructural Characteristics through Microstructurally Aware Augmentation of Scarce Data"** (*William F. Kent, Rochan Bajpai, Rachel C. Kurchin, William K. Epting, Harry W. Abernathy, Paul A. Salvador. Submitted 2026*). The methods are also described in the dissertation **Data Intensive Analysis of Solid Oxide Cell Microstructures** (*Doctoral dissertation, Carnegie Mellon University, 2025*). The datasets here are trained convolutional neural network (CNN) models for predicting key microstructural properties of solid oxide cell (SOC) electrodes from low-res, 2-channel 3D images, as well as some helpful code. The parameters for input images are provided in the paper. Sample data is provided in the file `Combined_anode_aug_dual_1k_examples` - that particular data was used to train `anode_all_aug.pth` and will work most accurately with that model. Please familiarize yourself with all caveats on accuracy and applicability, as detailed in the associated paper. Usage -------------------------- The basic usage is as follows, assuming `model_fn` is the path to the .pth file, and `X` is 2-channel input image(s) of the proper dimensions (either one image of shape `[2,12,24,24]`, or a batch of N input images of shape `[N,2,12,24,24]`): from CNN_inferencer import load_model_for_inference model = load_model_for_inference(model_fn) y_predicted = model(X) The model object automatically handles input scaling and output de-scaling based on the way the models were trained - in other words, pass in a 2-channel micro-CT image, and it will output microstructural property values in real units. ## Other model object attributes Note that model has useful attributes other than its forward pass model(X). * `model.output_descaler` - returns the output descaler object. Model does the de-scaling when generating inferences, but you may want to re-use this de-scaler on other values to e.g. compare predictions to ground truth from already-scaled training data. * `model.prop_names` - Gives the property names of the predicted y values, in order. Only exists if there's an output scaler as part of the model object, which there will be in the models provided here. ## Usage with sample data Here is a short script to use with the included sample data. from CNN_inferencer import display_predictions, load_model_for_inference, calculate_mape, parity_plot import h5py import numpy as np model_fn = 'anode_all_aug.pth' data_fn = 'Combined_anode_aug_dual_1k_examples.h5' N_samples = 200 figure_outdir = '.' model = load_model_for_inference(model_fn) with h5py.File(data_fn,'r') as f: XX = f['X'] #These are the 2-channel 3D images yy = f['y'] #These are the ground-truth microstructural properties, but they have been scaled for training - need to de-scale below N = XX.shape[0] #How many images total in the input data file #Run inferences on N_samples random samples from XX. #Run in a batch, much more efficient than one at a time. ii = np.random.choice(N,N_samples,replace=False) ii.sort() y_pred = model(XX[ii]) #Get the original/true (but normalized/scaled) values from the training dataset... #Because they were normalized, they are not in real units yet. So let's also de-scale them using model.output_scaler. y_true = model.output_scaler.transform(yy[ii]) #Let's display actual values for just 5 random ones for i in np.random.choice(N_samples,5,replace=False): display_predictions(y_true[i], y_pred[i], model.prop_names) #Make parity plots for each property (ground truth vs predicted values) #Also label each plot with the mean abs. percent error (MAPE) of the predicted values for i,key in enumerate(model.prop_names): mape = calculate_mape(y_true[:,i], y_pred[:,i]) parity_plot(y_true[:,i], y_pred[:,i], figure_outdir, key, extra_title=f' ({mape:.2f}% MAPE)')

3D microstructure↗

Effect of a Fine-Scale Layered Structure of the Atmosphere on Infrasound Signals from Fragmenting Meteoroids

We investigate the influence of a fine-scale (FS) layered structure in the atmosphere on the propagation of infrasound signals generated by fragmenting meteoroids. Using a pseudo-differential parabolic equation (PPE) approach, we model broadband acoustic signals from point sources at altitudes of 35–100 km. The presence of FS fluctuations in the stratosphere (37–45 km) and the lower thermosphere (100–120 km) modifies ray trajectories, causing multiple arrivals and prolonged signal durations at ground stations. In particular, meteoroids fragmenting at 80–100 km can produce two distinct thermospheric arrivals beyond 150 km range, while meteoroids descending to 50 km or below yield weak, long-lived arrivals within the acoustic shadow zone via antiguiding propagation and diffraction. Comparison with observed infrasound data confirms that FS-layered inhomogeneities can account for multi-arrival “N-waves,” broadening potential interpretations of meteoroid signals. The results also apply to other atmospheric-entry objects, such as sample return capsules, emphasizing how FS structure impacts shock wave propagation. In conclusion, our findings advance understanding of wavefield evolution in a layered atmosphere and have broad relevance for global infrasound monitoring of diverse phenomena (e.g., re-entry capsules, rocket launches, and large-scale explosions).

Aeroacoustics↗

Financial-technical co-design for capital-intensive, resource-responsive energy systems

Because of their capital-intensive operation, wind energy systems that are competitive in terms of the cost of the energy that they produce lead to risk-reward trade-offs that make their business cases less favorable than those of conventional energy generation technologies. However, wind energy systems tend to be designed to maximize energy production or minimize cost of energy rather than to maximize their business cases. In this work, we attempt to exploit designs specifically tailored to business cases. We develop a novel framework for analyzing energy systems that ties their design variables to monthly operating incomes using simple models and historical hourly market and resource data. Using this approach, we demonstrate that for a wind site with abundant wind resource in the California Independent System Operator market, we can control the trade-off between mean and 5th percentile monthly returns by choosing the specific power of the turbine at a fixed modeled initial capital cost. Our framework gives a measure of the risk-reward spectrum of energy generation assets that could be built at a given site with respect to the sub-annual resource/market variation.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Tula: Optimizing Time, Cost, and Generalization in Distributed Large-Batch Training

Distributed training increases the number of batches processed per iteration either by scaling-out (adding more nodes) or scaling-up (increasing the batch-size). However, the largest configuration does not necessarily yield the best performance. Horizontal scaling introduces additional communication overhead, while vertical scaling is constrained by computation cost and device memory limits. Thus, simply increasing the batch-size leads to diminishing returns: training time and cost decrease initially but eventually plateaus, creating a knee-point in the time/cost vs. batch-size pareto curve. The optimal batch-size therefore depends on the underlying model, data and available compute resources. Large batches also suffer from worse model quality due to the well-known “generalization gap”. In this paper, we present Tula, an online service that automatically optimizes time, cost, and convergence quality for large-batch training of convolutional models. It combines parallel-systems modeling with statistical performance prediction to identify the optimal batchsize. Tula predicts training time and cost within 7.5−14% error across multiple models, and achieves up to 20× overall speedup and improves test accuracy by ≈9% on average over standard large-batch training on various vision tasks, thus successfully mitigating the generalization gap and accelerating training at the same time.

Tyagi, Sahil [ORNL] (ORCID:0009000783144745)↗

Reducing AI RAG Hallucination by Optimizing Routing Techniques

Large Language Models (LLMs), such as ChatGPT, tend to “hallucinate”, meaning they confidently generate false information. Retrieval Augmented Generation (RAG) attempts to diminish hallucination by providing context to the LLM from data stores (indexes) containing relevant information. The LLM uses this context to formulate its response. RAG systems can still suffer from hallucination because of bad embeddings or ineffective routing. For example, a router will often return context from an irrelevant index, resulting in a hallucinated answer. In this study, we aim to minimize the frequency of routing hallucinations by optimizing Index Summary Routing.

97 MATHEMATICS AND COMPUTING↗

Development of a Test-Bed for Testing and Refining EarthEn’s Supercritical CO 2 Based Energy Storage System

EarthEn’s energy storage concept leverages supercritical carbon dioxide (sCO 2 ) as a working fluid and relies on compact, high-performance components operating at elevated pressures and temperatures. To accelerate component development and reduce technical risk prior to larger-scale demonstrations, Oak Ridge National Laboratory (ORNL) developed a 100 kW-scale sCO 2 test-bed under a Cooperative Research and Development Agreement with EarthEn (CRADA NO. NFE-24-10050). The objective of the work was to design and construct a flexible experimental facility capable of reproducing key thermodynamic state points and heat-transfer conditions relevant to EarthEn’s thermal energy storage (TES) cycle, with particular emphasis on enabling development and evaluation of next-generation heat exchangers and TES concepts. The test-bed consists of a closed-loop sCO 2 circulation system housed within an open-topped enclosure. In its as-installed configuration, dense-phase sCO 2 is recirculated through a printed circuit recuperator, an electrically heated section, a throttling device used to simulate turbine expansion, and a water-cooled printed circuit heat exchanger that rejects heat to the building chilled-water system before returning to the pump. The pump is driven by a variable frequency drive, enabling controlled adjustment of flow and operating point. A comprehensive instrumentation suite was integrated to support both safe operation and high-quality data collection. Installed sensors include Coriolis flow meters for sCO 2 flow rate and density, resistance temperature detectors and thermocouples distributed throughout the loop (including the heated section and key heat exchanger ports), and pressure transducers for absolute and differential pressure measurements. The facility was designed to support high-pressure (19 MPa nominal) and high-temperature (575°C nominal) operation with credited overpressure protection provided by a rupture disk. Nominal operating conditions were selected to support 100 kW-class testing while maintaining flexibility for non-heated and heated shakedown, control development, and future integration of advanced TES test sections. In parallel with facility development, a system-level thermal-hydraulic model was created using Modelica-based tools to support component sizing, anticipate performance over targeted test conditions, and establish a framework for future model calibration against experimental data. At the conclusion of the project performance period, the facility was in final assembly, and the pressure boundary was nearly completed. However, several practical challenges associated with high-pressure/high-temperature systems and specialized component procurement impacted schedule and prevented initial pump-driven operation and full commissioning within the available resources. This report documents the as-built design, operating capabilities, and instrumentation, and it summarizes key lessons learned related to heater fabrication and testing, first-of-a-kind assembly factors, specialty flange supply constraints, and fill pump corrective actions. Finally, it outlines a phased plan for future commissioning and experimental campaigns, including control and instrumentation shakedown, heater characterization, model calibration, and testing at state points representative of EarthEn’s TES cycle.

25 ENERGY STORAGE↗

(Doublon) Benchmarking of Different Inverse Point Kinetics Implementations for an Autocorrected Reactimeter Algorithm

In November 2017, the Transient Reactor Test Facility returned to operation. Since that time, many transient test series have been completed, such as the Transient Heatsink Overpower Response capsule (THOR), the Transient Water Irradiation System for TREAT (TWIST), and Sirius. Each has provided valuable data for materials performance and reactor safety that can be applied in future designs. During each experimental series, detector count rates provided important information on the core behavior during transients. However, a limitation of these data is that variations in the neutron distribution during experiments can cause errors when attempting to infer reactivity evolution from detector signals. Neutron physics codes can be used to compute the flux shape variations. However, this is a poor solution when the experimental data is used for code verification, validation and uncertainty quantification. Indeed, if the output of the code is used both as a reference and to correct what the reference is compared to, the circular dependency limits the quality of the verification, validation and uncertainty quantification approach. To overcome this problem, the autocorrected reactimeter algorithm (ACRA) has been developed. This approach infers a time-dependent reactivity evolution by testing different spatial corrections and selecting the one that minimizes reactivity variations when the core is in a frozen configuration (i.e., when there is no variation in parameters affecting reactivity). However, the scope of this method was limited to transients where there were negligible thermal feedback. Indeed, the core is never in a frozen configuration when the fuel temperature varies during the whole transient. This is our motivation for developing an improved version of the ACRA that does not require frozen configurations. To develop this new algorithm, we need a precise and unbiased implementation of the inverse point kinetic equations (IPKEs) as any error in the reactivity evaluation will be propagated into the choice of the optimal spatial correction. Indeed, the previous reactimeter algorithm would use approximations, such as a negligible flux amplitude derivative, to focus on rapidity. For the numerical validation of ACRA, we aim at absolute error under for reactivity derived from signals similar to the one of this study. In this summary, we test eight different IPKE implementations. Each will process a mockup signal built for this study, similar to those that the future ACRA will process. Each reactivity output will be compared to the reference reactivity that has been used to generate the mockup signal. The implementation minimizing the difference with the reference reactivity will be used in the development of a new ACRA formulation.

73 - NUCLEAR PHYSICS AND RADIATION PHYSICS↗

AmeriFlux US-EKH Elkhorn Slough Hester Marsh

This is the AmeriFlux version of the carbon flux data for the site US-EKH Elkhorn Slough Hester Marsh. Site Description - Salt marsh restored in 2018 by adding soil to raise 61 acres of former salt marsh to an elevation that allows marsh plants to return and keep pace with projected sea level rise. Marsh is mostly bare soil and tidal channels. Marsh is flooded and drained diurnally.

Paytan, Adina [University of California, Santa Cru↗

Leveraging Large Language Models for Real-World Data Evidence: A Framework for Automated Treatment Extraction and Data Harmonization

Background: The ability to comprehensively collect treatment information from cancer patient medical records would enable studies to evaluate real-world benefits and risks tied to specific treatments. Currently, it is difficult to system- atically collect high-quality treatment information because it is often stored in unstructured text. Manually extracting and standardizing drug and regimen data is time-intensive. Recent advances in large language models (LLMs) offer a potential solution for automated extraction of structured treatment information from clinical text. Objective: This study systematically evaluates the utility of four LLMs from the Llama family for automated extraction of oncology treatment information from clinical text. This information can guide researchers using cancer registry data to provide insights into cancer care and outcomes beyond clinical trials. Methods: Four instruction-tuned Llama models with varying parameter counts (1B, 3B, 8B, and 70B) were evaluated for their ability to extract treatment information from clinical documents. A unified oncology knowledge base integrating seven major public data sources was developed to standardize and normalize extracted entities—a critical step for harmonizing data from diverse sources. Extracted treatment data were compared against expert-annotated ground truth. Model performance was assessed using accuracy metrics (Precision, Recall, F1-Score) and opera- tional feasibility metrics, including processing speed and structural compliance of the output. Results: A strong positive correlation was observed between model size and extraction accuracy. F1-score improved from 0.609 for the 1B model to 0.710 (3B), 0.807 (8B), and 0.828 (70B). While larger models demonstrated superior accuracy and compliance, they incurred higher computational costs. The modest performance difference between 8B and 70B suggests diminishing returns with increasing model size. Conclusions: LLMs represent a viable technology for automating oncology treatment extraction. The 8B-parameter model emerged as a highly effective option, balancing high accuracy and computational efficiency. Selecting an appropriate LLM for deployment in cancer registries involves a trade-off between desired accuracy and available operational resources. Harmonizing extracted entities with the oncology knowledge base facilitates standardized integration into common data models, enhancing data quality for real-world evidence analyses.

artificial intelligence↗