Search NASA⌕ Search

SEARCH · Search NASA

Results for “error analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 559 records · Page 31

RELAP5-3D validation studies based on the High Temperature Test facility

In the spring and summer of 2019, experiments were conducted at the High Temperature Test Facility (HTTF) that form the basis of an upcoming high-temperature gas-cooled reactor (HTGR) thermal hydraulics (T/H) benchmark. HTTF is an integral effects test facility for HTGR T/H modeling validation. This paper presents RELAP5-3D models of two of those experiments: PG-27, a pressurized conduction cooldown (PCC); and PG-29, a depressurized conduction cooldown (DCC). These models used the RELAP5-3D model of HTTF originally developed by Paul Bayless as a starting point. The sensitivity analysis and uncertainty quantification code, RAVEN was used to perform calibration studies for the steady-state portion of PG-27. Here we developed four PG-27 calibrations based on steady-state conditions. These calibrations all used an effective thermal conductivity equal to 36 % of the measured thermal conductivity, but they differed with respect to the frictional pressure drops and radial conduction models. These models all captured the trends in steady-state temperature distributions and transient temperature behavior well. All four calibrations show room for improvement in predicting the transient temperature rise. The smallest error in temperature rise during the transient was a 21 % underprediction, and the largest was a 48 % underprediction. The errors in transient temperature rise are largely a result of a mismatch in power density between the RELAP5-3D model and the experiment due to the location of active heater rods along the boundary between heat structures in the model. The best of these calibrations was applied to PG-29 to model the DCC. Once again, temperatures during the transient were underpredicted but trends in temperature were captured. The RELAP5-3D model captured trends in the data but could not reproduce measured temperatures exactly. This result is not attributed to deficiencies in the experimental data or to RELAP5–3D itself. Rather, this result likely arises due to the some of the assumptions and decisions made when the RELAP5-3D model was first developed, prior to the execution of HTTF experiments. An agreement in prediction of temperature trends but challenges reproducing HTTF temperatures within measurement uncertainty is consistent with previous analyses of HTTF in the literature. Future RELAP5-3D validation activities centered around HTTF may be able to provide greater insight into the code’s capabilities for HTGR modeling with a more finely nodalized model.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Remote Instrumentation and Data Acquisition

This poster outlines the development and implementation of a remote data acquisition system for waveform analysis using a Rohde & Schwarz oscilloscope. The project involved capturing waveform data, and transferring it to a local machine for visualization and analysis. The core logic was developed in C++ with a focus on object oriented programming and the use of polymorphism so the main application can interact with any instrument without knowing its exact type, simplifying the overall logic and making it easier to add or swap out components without changing the rest of the codebase.. The system issues Standard Commands for Programmable Instruments (SCPI) via a socket connection and parses the oscilloscope s ASCII waveform data. The C++ application was containerized using Docker for ease of portability, and reproducibility. Emphasis was placed on secure networking practices, error handling, and effective data capture. The report describes the technical steps taken, challenges encountered, and future work, providing insight into the practical integration of hardware interfacing with remote computational environments.

Parikh, Jaymil [Illinois U., Urbana]↗

Remote Instrumentation and Data Acquisition: An Internship Research Report

This report outlines the development and implementation of a remote data acquisition system for waveform analysis using a Rohde & Schwarz oscilloscope. The project involved capturing waveform data, and transferring it to a local machine for visualization and analysis. The core logic was developed in C++ with a focus on object oriented programming and the use of polymorphism so the main application can interact with any instrument without knowing its exact type, simplifying the overall logic and making it easier to add or swap out components without changing the rest of the codebase.. The system issues Standard Commands for Programmable Instruments (SCPI) via a socket connection and parses the oscilloscope’s ASCII waveform data. The C++ application was containerized using Docker for ease of portability, and reproducibility. Emphasis was placed on secure networking practices, error handling, and effective data capture. The report describes the technical steps taken, challenges encountered, and lessons learned, providing insight into the practical integration of hardware interfacing with remote computational environments.

Parikh, Jaymil [Fermilab]↗

Transfer learning for analysis of collective and non-collective Thomson scattering spectra

Thomson scattering (TS) diagnostics provide reliable, minimally perturbative measurements of fundamental plasma parameters, such as electron density (⁠n e ) and electron temperature (⁠T e ⁠). Deep neural networks can provide accurate estimates of ⁠n e and T e when conventional fitting algorithms may fail, such as when TS spectra are dominated by noise, or when fast analysis is required for real-time operation. Although deep neural networks typically require large training sets, transfer learning can improve model performance on a target task with limited data by leveraging pre-trained models from related source tasks, where select hidden layers are further trained using target data. We present five architecturally diverse deep neural networks, pre-trained on synthetic TS data and adapted for experimentally measured TS data, to evaluate the efficacy of transfer learning in estimating n e and T e in both the collective and non-collective scattering regimes. We evaluate errors in n e and T e estimates as a function of training set size for models trained with and without transfer learning, and we observe decreases in model error from transfer learning when the training set contains ≲ 200 experimentally measured spectra.

Artificial neural networks↗

Deep-ultraviolet ptychographic pocket-scope (DART): mesoscale lensless molecular imaging with label-free spectroscopic contrast

The mesoscale characterization of biological specimens has traditionally required compromises between resolution, field-of-view, depth-of-field, and molecular specificity, with most approaches relying on external labels. Here we present the Deep-ultrAviolet ptychogRaphic pockeT-scope (DART), a handheld platform that transforms label-free molecular imaging through intrinsic deep-ultraviolet spectroscopic contrast. By leveraging biomolecules’ natural absorption fingerprints and combining them with lensless ptychographic microscopy, DART resolves down to 308-nm linewidths across centimeter-scale areas while maintaining millimeter-scale depth-of-field. The system’s virtual error-bin methodology effectively eliminates artifacts from limited temporal coherence and other optical imperfections, enabling high-fidelity molecular imaging without lenses. Through differential spectroscopic imaging at deep-ultraviolet wavelengths, DART quantitatively maps nucleic acid and protein distributions with femtogram sensitivity, providing an intrinsic basis for explainable virtual staining. We demonstrate DART’s capabilities through imaging of tissue sections, cytopathology specimens, blood cells, and neural populations, revealing detailed molecular contrast without external labels. The combination of high-resolution molecular mapping and broad mesoscale imaging in a portable platform opens new possibilities from rapid clinical diagnostics, tissue analysis, to biological characterization in space exploration.

60 APPLIED LIFE SCIENCES↗

Modal Field Reconstruction in Resonant Cavities in the Fundamental and Undermoded Frequency Regimes

Theory, simulations, and experiments are presented that demonstrate reconstruction of electromagnetic fields in a cavity from sparse probe measurements. Such techniques are often referred to as virtual sensing, allowing fields at unobserved locations to be predicted. These methods are appropriate for the fundamental and undermoded regimes, providing the ability to estimate fields (and shielding effectiveness) throughout an arbitrarily shaped cavity from a few judiciously spaced probes. A modal simulation method is implemented that allows the response of arbitrarily shaped cavities to be rapidly computed with respect to varying probe locations and slot parameters, enabling statistical analysis of probe placement on reconstruction performance. A cylindrical vessel with numerous probe holes is developed for experiments, referred to as Perforated Vessel 2 (PV2). Experiments are performed on the vessel with and without a steel box inside, where transmit power is delivered into the vessel either through probes (probe injection) or through slots using an external antenna (slot excitation). Simulations and experiments illustrate that when the number of probes is minimal (equal to the number of mode coefficients to be estimated at each frequency), probe placement is critical to avoid missed peaks and to have acceptable reconstruction error. Probe placement becomes less important as the number of probes is increased, but care is still required to avoid probe locations giving poor performance.

42 ENGINEERING↗

Benchmarking Correlation-Consistent Basis Sets for Frequency-Dependent Polarizabilities with Multiresolution Analysis

This paper presents the first converged frequency-dependent HF polarizability results for general molecules, on a set of 89 closed-shell atoms and molecules. The solver employs multiresolution analysis (MRA) in a multiwavelet basis to compute both ground and response states to a guaranteed precision, which are validated against independent numerical grid calculations on atoms and linear molecules. The MRA ground-state energies and response properties are used to evaluate results in correlation-consistent basis sets up to 5Z augmented with either single or double diffuse functions and core-polarization functions. Systematic trends are revealed through consideration of chemical composition as well as the use of machine learning to cluster convergence trends, the latter suggesting the possibility of learning and correcting basis-set error.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Towards an Introspective Dynamic Model of Globally Distributed Computing Infrastructures

Large-scale scientific collaborations like ATLAS, Belle II, CMS, DUNE, and others involve hundreds of research institutes and thousands of researchers spread across the globe. These experiments generate petabytes of data, with volumes soon expected to reach exabytes. Consequently, there is a growing need for computation, including structured data processing from raw data to consumer-ready derived data, extensive Monte Carlo simulation campaigns, and a wide range of end-user analysis. To manage these computational and storage demands, centralized workflow and data management systems are implemented. However, decisions regarding data placement and payload allocation are often made disjointly and via heuristic means. A significant obstacle in adopting more effective heuristic or AI-driven solutions is the absence of a quick and reliable introspective dynamic model to evaluate and refine alternative approaches. In this study, we aim to develop such an interactive system using real-world data. By examining job execution records from the PanDA workflow management system, we have pinpointed key performance indicators such as queuing time, error rate, and the extent of remote data access. The dataset includes five months of activity. Additionally, we are creating a generative AI model to simulate time series of payloads, which incorporate visible features like category, event count, and submitting group, as well as hidden features like the total computational load—derived from existing PanDA records and computing site capabilities. These hidden features, which are not visible to job allocators, whether heuristic or AI-driven, influence factors such as queuing times and data movement.

kilic, Ozgur Ozan [Brookhaven National Laboratory ↗

Rotated-Droop Control for Enhanced Stability and Power Decoupling in Microgrids With Complex Line Impedances

Classical droop control in microgrids predominantly assumes inductive line impedance, which simplifies implementation but causes power coupling and steady-state errors in systems with resistive and inductive lines. This paper proposes a rotated-droop control strategy for grid-forming inverters that reformulates the power equations by incorporating the magnitude and angle of the line impedance within a rotated reference frame. This method enhances the decoupling of active and reactive power dynamics without increasing complexity or requiring communication links. A small-signal state-space model was created to capture the dynamic behavior of the system under varying impedance parameters, preserving the original droop gains by rotating the power control structure. This allows eigenvalue-based stability analysis and enhances damping and transient performance. Simulation and experimental results validated the improved power-sharing performance, faster response, and robustness of the proposed method under different impedance conditions. This approach maintains the decentralized structure of the conventional droop control while enabling greater adaptability and scalability, making it suitable for modern inverter-based microgrids with dynamic topologies.

Campo-Ossa, Daniel Dario [Univ. of Puerto Rico, Ag↗

Heavy-Duty Nonroad Material Handler Electrification Part 1: Real-World Drive Cycle Development

Knowing a detailed operating cycle is critical for developing and testing equipment. Operating cycles can be separated by two clear distinctions: (1) regulatory or non-regulatory and (2) application at the engine-only or full machine level. The Environmental Protection Agency’s (EPA) Nonroad Transient Cycle (NRTC) may be a good representation of engine use in many types of equipment, but there is a gap in standardized and validated drive cycles specifically for nonroad material handlers. Lacking a standardized drive cycle makes it difficult to accurately benchmark machine performance and validate new powertrain technologies. The objective of this investigation is to illustrate the development of a custom drive cycle augmented with real-world customer use data that serves multiple purposes: (1) understand the range of operation and utilization that formulated inputs for electrified architecture analysis and (2) develop a repetitive and consistent maneuver to establish baseline energy consumption enabling equivalent comparison to future electrified prototype builds. This article presents a solution specifically for a 23-ton nonroad material handler in which material handling, machine transport, and extended idle were homologated to form representative short cycles defined by machine velocity and hydraulic cylinder position. The most intensive material handling short cycles had a load factor of 40% and an average fuel rate of 16 L/h. Combined with a visual aid, the short cycles exhibited low variability, having less than 5% root mean square (RMS) error in lift and reach position with respect to the average. The machine’s performance on these short cycles at the Advanced Power Systems Research Center (APSRC) was compared to results from two real-world customer locations operating the instrumented test machine in a cyclical manner, and for similar ground conditions were found to be comparable in fuel consumption.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Equilibrium Core Model for Micro Pebble Bed Reactors Using OpenMC

Estimating the equilibrium state for pebble bed reactors (PBRs) presents complex challenges as it requires simultaneous consideration of changes in the pebbles’ movement as well as their fuel compositions. Whereas traditional approaches use multigroup diffusion codes for neutronics calculations of PBRs’ equilibrium state, the double-heterogeneity of PBRs complicates neutron cross-section generation. Continuous-energy Monte Carlo (MC) methods are better suited for detailed PBR analysis because of their natural handling of double-heterogeneity, but they demand substantially more computational resources. Here, this study introduces a novel method for efficiently estimating the equilibrium state in small and micro PBRs with reduced computational cost. The method is anticipated to accelerate the processes of core design and performing parametric studies for utilizing advanced fuel and structural materials. The HTR-10 reactor design was used for validating the method’s predictions and evaluating its computational efficiency. When compared to reference calculation values from the literature, criticality (k-effective) was predicted to be approximately within the margin of error of the MC transport calculation, average core power density (in megawatts per cubic meter) was predicted within 2.5% relative error, and maximum thermal flux (10 13 n/cm 2 .s −1 ) was predicted within 1.8% relative error. The calculated inventory of fission products and fuel composition in the equilibrium core were within 15% and 16.6%, respectively, when compared to reported values from the literature. The difference is attributed to variance in the considered values of the core temperature, which was found to significantly affect the depletion analyses.

Equilibrium core↗

Robust error calibration for serial crystallography

Serial crystallography is an important technique with unique abilities to resolve enzymatic transition states, minimize radiation damage to sensitive metalloenzymes and perform de novo structure determination from micrometre-sized crystals. This technique requires the merging of data from thousands of crystals, making manual identification of errant crystals unfeasible. cctbx.xfel.merge uses filtering to remove problematic data. However, this process is imperfect, and data reduction must be robust to outliers. We add robustness to cctbx.xfel.merge at the step of uncertainty determination for reflection intensities. This step is a critical point for robustness because it is the first step where the data sets are considered as a whole, as opposed to individual lattices. Robustness is conferred by reformulating the error-calibration procedure to have fewer and less stringent statistical assumptions and incorporating the ability to down-weight low-quality lattices. We then apply this method to five macromolecular XFEL data sets and observe the improvements to each. The appropriateness of the intensity uncertainties is demonstrated through internal consistency. This is performed through theoretical CC 1/2 and I /σ relationships and by weighted second moments, which use Wilson's prior to connect intensity uncertainties with their expected distribution. This work presents new mathematical tools to analyze intensity statistics and demonstrates their effectiveness through the often underappreciated process of uncertainty analysis.

Mittan-Moreau, David W.↗

Characteristics of the IBEX Ribbon and Their Implications for a Source Region Outside the Heliopause

This paper presents a comprehensive exploration of the Interstellar Boundary Explorer energetic neutral atom (ENA) ribbon, focusing on its spatial and temporal variations over 14 yr. Methodological advancements, including a refined map modeling procedure and a new ribbon separation technique with appropriate error propagation, enable a detailed investigation of the ribbon’s features. Utilizing statistically robust metrics, this study reveals details of the ribbon across energy and time. Key findings include energy- and time-dependent variations in flux, angular radius, ribbon profile width, and higher moments. By applying these metrics, we reveal new complexity to the evolution of the ribbon over time, highlighting the nuanced relationship between it and the solar wind. Furthermore, the study examines for the first time the ribbon as it passes through the starboard/heliotail region (Lon EC 120°–180°), revealing properties distinct from other portions of the ribbon. The analysis uncovers an anticorrelation between ribbon width and flux, which provides quantitative support for a multisource ribbon created by a combination of solar wind neutrals that generate a spatiall narrow ribbon component and heliosheath neutrals giving rise to a broad component. Finally, differences in the temporal evolution of the ENA flux at different energies provide additional support that the location of the ribbon source region is beyond the heliopause.

79 ASTRONOMY AND ASTROPHYSICS↗

High-n Rydberg transition spectroscopy for heavy impurity transport studies in W7-X (invited)

Here, we present a novel spectroscopy approach to investigate impurity transport by analyzing line-radiation following high-n Rydberg transitions. While high-n Rydberg states of impurity ions are unlikely to be populated via impact excitation, they can be accessed by charge exchange (CX) reactions along the neutral beams in high-temperature plasmas. Hence, localized radiation of highly ionized impurities, free of passive contributions, can be observed at multiple wavelengths in the visible range. For the analysis and modeling of the observed Rydberg transitions, a technique for calculating effective emission coefficients is presented that can well reproduce the energy dependence seen in datasets available on the OPEN-ADAS database. By using the rate coefficients and comparing modeling results with the new high-n Rydberg CX measurements, impurity transport coefficients are determined with well-documented 2σ confidence intervals for the first time. This demonstrates that high-n Rydberg spectroscopy provides important constraints on the determination of impurity transport coefficients. By additionally considering Bolometer measurements, which provide constraints on the overall impurity emissivity and, therefore, impurity densities, error bars can be reduced even further.

Instruments & Instrumentation↗

Convergent Manufacturing of Large-Scale Components for Nuclear Applications, via Additive Manufacturing and Powder Metallurgy Hot Isostatic Pressing

Powder metallurgy (PM)–hot isostatic pressing (PM-HIP) has long been recognized as a powerful route for producing fully dense, near net shape metallic components. By consolidating powders under high temperature and pressure, HIP provides isotropic properties, uniform microstructures, and scalability to complex geometries that are vital for sectors such as aerospace, energy, and nuclear power. Yet despite these advantages, the technology has remained constrained by costly trial and error canister fabrication, limitations of conventional forging, and incomplete knowledge about how the canister design influences final part properties. Additive manufacturing (AM), by contrast, thrives on design freedom and geometric flexibility but struggles with speed, scalability, and cost when applied to very large structures. The research presented in this report investigated how a convergent manufacturing approach, combining AM with PM-HIP, can merge the strengths of both technologies, leveraging AM’s flexibility for canister design and HIP’s consolidation capability to deliver reliable, large, and complex parts. The work progressed through three case studies that built on one another in scale and complexity. Small cylindrical canisters fabricated by conventional methods, laser powder bed fusion, and directed energy deposition were filled with stainless steel powders and subjected to HIP. The resulting parts demonstrated near-full density and mechanical properties on par with wrought stainless steel, showing for the first time that AM canisters can be a direct substitute for conventional ones without sacrificing quality. The next step involved a medium-scale, noncentrosymmetric T-valve, which is an enclosed, multibranch geometry that tested the limits of AM + PM-HIP integration. The T-valve achieved predictable shrinkage and uniform densification, confirming feasibility for enclosed designs. However, this study also revealed oxide inclusions and interfacial challenges at the AM + HIP boundary, underscoring the critical importance of controlling interface chemistry and employing robust, in situ strategies, such as melt pool monitoring and thermal monitoring, coupled with nondestructive evaluation techniques such as x-ray computed tomography. Finally, the effort culminated in fabricating a large-scale impeller weighing nearly 2000 lb and spanning 5 ft in diameter. Produced via multirobot wire arc AM and hot isostatic pressed to near-full density, the impeller validated industrial-scale feasibility. Predictive models closely matched experimental shrinkage, tensile properties were spatially uniform across the component, and the AM + PM-HIP interface proved mechanically sound despite the presence of oxide-decorated prior particle boundaries. This large-scale demonstration is a major milestone, showing that hybrid AM + PM‑HIP can reliably deliver components at reactor-relevant scales. Collectively, these studies charted a logical pathway: small-scale work built scientific confidence, medium-scale work highlighted opportunities and challenges, and large-scale work proved industrial impact. The overarching conclusion of this report is that AM + PM-HIP should not be seen as a replacement for forging but as a complementary pathway that provides the US with flexibility, resilience, and new options for manufacturing nuclear-grade components. Looking ahead, several directions emerge as critical to sustaining progress. Predictive modeling must become faster, more accessible, and more accurate, with digital twins and machine learning reducing reliance on trial and error. Powders and alloys must be optimized for HIP, with improved cleanliness, reduced oxides, and tailored chemistries that enhance creep, fatigue, and irradiation resistance. Interfaces between AM and HIP regions must be better engineered through coatings, machining strategies, and surface treatments to mitigate oxide formation and ensure reliable bonding to explore opportunities for HIP of targeted compositional parts, as well as multimaterial HIP cladding applications. Monitoring and nondestructive evaluation need to expand, incorporating multimodal sensors, x-ray computed tomography, and real-time data integration through platforms such as Pelican. At the same time, the pathway to industrial adoption requires techno-economic analysis, machinability studies, and qualification frameworks aligned with industry and regulatory standards. Finally, workforce and academic engagement must be strengthened. Programs that train technicians and engineers for US Navy and US Department of Energy manufacturing challenges should be paired with academic partnerships to support fundamental research, with open sharing of non-export-controlled data to accelerate innovation and build the next generation of experts. In conclusion, this report demonstrates that hybrid AM + PM-HIP is scientifically viable and strategically important. By combining the design agility of AM with the consolidation strength of HIP and embedding modeling, monitoring, and workforce development, this approach provided a transformative new capability for US manufacturing. The path forward is clear: hybrid AM + PM-HIP is not just a promising research direction but is also potentially an industrially relevant pathway that can reshape how nuclear-grade components are designed, qualified, and deployed.

36 MATERIALS SCIENCE↗

Unveiling the transferability of PLSR models for leaf trait estimation: lessons from a comprehensive analysis with a novel global dataset

Leaf traits are essential for understanding many physiological and ecological processes. Partial least squares regression (PLSR) models with leaf spectroscopy are widely applied for trait estimation, but their transferability across space, time, and plant functional types (PFTs) remains unclear. We compiled a novel dataset of paired leaf traits and spectra, with 47 393 records for >700 species and eight PFTs at 101 globally distributed locations across multiple seasons. Using this dataset, we conducted an unprecedented comprehensive analysis to assess the transferability of PLSR models in estimating leaf traits. While PLSR models demonstrate commendable performance in predicting chlorophyll content, carotenoid, leaf water, and leaf mass per area prediction within their training data space, their efficacy diminishes when extrapolating to new contexts. Specifically, extrapolating to locations, seasons, and PFTs beyond the training data leads to reduced R 2 (0.12–0.49, 0.15–0.42, and 0.25–0.56) and increased NRMSE (3.58–18.24%, 6.27–11.55%, and 7.0–33.12%) compared with nonspatial random cross-validation. The results underscore the importance of incorporating greater spectral diversity in model training to boost its transferability. These findings highlight potential errors in estimating leaf traits across large spatial domains, diverse PFTs, and time due to biased validation schemes, and provide guidance for future field sampling strategies and remote sensing applications.

59 BASIC BIOLOGICAL SCIENCES↗

Emerging low-cloud feedback and adjustment in global satellite observations

From mid-2003 to mid-2024, a global decrease in low-cloud amount enhanced the absorption of solar radiation by 0.22±0.07 W m −2 per decade (±1σ range), accelerating the energy imbalance trend during that period (0.44 W m −2 per decade). Through controlling factor analysis, here we show that the low-cloud trend is due to a combination of cloud feedback and adjustments to greenhouse gases and aerosols (respectively 0.09±0.02, 0.05±0.03, and 0.03±0.03 W m −2 per decade), which jointly account for 74 % of the trend. The contribution of natural climate variability is weak but uncertain (0.01±0.08 W m −2 per decade), owing to a poorly constrained trend in boundary-layer inversion strength. Importantly, the observed low-cloud radiative trend lies well within the range of values simulated by contemporary global climate models under conditions close to present day. Any systematic model error in the representation of present-day global energy imbalance trends is thus likely to originate in processes unrelated to low clouds.

Geosciences↗

Can Large Language Models Understand Intermediate Representations?

Intermediate Representations (IRs) are essential in compiler design and program analysis, yet their comprehension by Large Language Models (LLMs) remains underexplored. This paper presents a pioneering empirical study to investigate the capabilities of LLMs, including GPT-4, GPT-3, Gemma 2, LLaMA 3.1, and Code Llama, in understanding IRs. We analyze their performance across four tasks: Control Flow Graph (CFG) reconstruction, decompilation, code summarization, and execution reasoning. Our results indicate that while LLMs demonstrate competence in parsing IR syntax and recognizing high-level structures, they struggle with control flow reasoning, execution semantics, and loop handling. Specifically, they often misinterpret branching instructions, omit critical IR operations, and rely on heuristic-based reasoning, leading to errors in CFG reconstruction, IR decompilation, and execution reasoning. The study underscores the necessity for IR-specific enhancements in LLMs, recommending fine-tuning on structured IR datasets and integration of explicit control flow models to augment their comprehension and handling of IR-related tasks.

Jiang, Hailong↗