Search NASA⌕ Search

SEARCH · Search NASA

Results for “Domain decomposition”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 271 records · Page 15

A Comparison of PETSC Library and HPF Implementations of an Archetypal PDE Computation

Two paradigms for distributed-memory parallel computation that free the application programmer from the details of message passing are compared for an archetypal structured scientific computation a nonlinear, structured-grid partial differential equation boundary value problem using the same algorithm on the same hardware. Both paradigms, parallel libraries represented by Argonne's PETSC, and parallel languages represented by the Portland Group's HPF, are found to be easy to use for this problem class, and both are reasonably effective in exploiting concurrency after a short learning curve. The level of involvement required by the application programmer under either paradigm includes specification of the data partitioning (corresponding to a geometrically simple decomposition of the domain of the PDE). Programming in SPAM style for the PETSC library requires writing the routines that discretize the PDE and its Jacobian, managing subdomain-to-processor mappings (affine global- to-local index mappings), and interfacing to library solver routines. Programming for HPF requires a complete sequential implementation of the same algorithm, introducing concurrency through subdomain blocking (an effort similar to the index mapping), and modest experimentation with rewriting loops to elucidate to the compiler the latent concurrency. Correctness and scalability are cross-validated on up to 32 nodes of an IBM SP2.

Hayder, M. Ehtesham↗

Phase-Field Methods for Structure Evolution in Sheared Multiphase Systems

A homogeneous disordered phase separates into ordered structures when quenched into a broken-symmetry phase. The competition of broken-symmetry phases to select an equilibrium state may be studied in terms of coarse-grained order parameters described by a suitable Landau free-energy function. A network of equilibrium-phase domains develops on quenching and coarsens with time with a topology that may be controlled by shear. We use three-dimensional simulations, in which time-dependent models for conserved-order parameters coupled to Navier-Stokes fluid models are solved, to investigate the evolution of such domains, e.g. spinodal decompositions of polymeric materials under shear. The numerical problems are formidable because of the strong nonlinearities inherent in the coupled model, and these are amongst the first 3D calculations undertaken. In linear shear fields we find stable nanostrings, also recently seen in experiments. The affinity of the ordered phases to boundaries plays a role in the form of the structures that develop, with stacked plate-like phase distributions emerging under certain conditions. Such methods appear quite promising for design and analysis of multiphase and complex fluid formulations. The behavior of foams in such conditions is of particular interest in microgravity environments. Additional information can be found in the original extended abstract.

Badalassi, Vittorio↗

The Planning Execution Monitoring Architecture

The Planning Execution Monitoring (PEM) architecture is a design concept for developing autonomous cockpit command and control software. The PEM architecture is designed to reduce the operations costs in the space transportation system through the use of automation while improving safety and operability of the system. Specifically, the PEM autonomous framework enables automatic performance of many vehicle operations that would typically be performed by a human. Also, this framework supports varying levels of autonomous control, ranging from fully automatic to fully manual control. The PEM autonomous framework interfaces with the core flight software to perform flight procedures. It can either assist human operators in performing procedures or autonomously execute routine cockpit procedures based on the operational context. Most importantly, the PEM autonomous framework promotes and simplifies the capture, verification, and validation of the flight operations knowledge. Through a hierarchical decomposition of the domain knowledge, the vehicle command and control capabilities are divided into manageable functional "chunks" that can be captured and verified separately. These functional units, each of which has the responsibility to manage part of the vehicle command and control, are modular, re-usable, and extensible. Also, the functional units are self-contained and have the ability to plan and execute the necessary steps for accomplishing a task based upon the current mission state and available resources. The PEM architecture has potential for application outside the realm of spaceflight, including management of complex industrial processes, nuclear control, and control of complex vehicles such as submarines or unmanned air vehicles.

Wang, Lui↗

Approximation, abstraction and decomposition in search and optimization

In this paper, I discuss four different areas of my research. One portion of my research has focused on automatic synthesis of search control heuristics for constraint satisfaction problems (CSPs). I have developed techniques for automatically synthesizing two types of heuristics for CSPs: Filtering functions are used to remove portions of a search space from consideration. Another portion of my research is focused on automatic synthesis of hierarchic algorithms for solving constraint satisfaction problems (CSPs). I have developed a technique for constructing hierarchic problem solvers based on numeric interval algebra. Another portion of my research is focused on automatic decomposition of design optimization problems. We are using the design of racing yacht hulls as a testbed domain for this research. Decomposition is especially important in the design of complex physical shapes such as yacht hulls. Another portion of my research is focused on intelligent model selection in design optimization. The model selection problem results from the difficulty of using exact models to analyze the performance of candidate designs.

Ellman, Thomas↗

An analysis of scatter decomposition

A formal analysis of a powerful mapping technique known as scatter decomposition is presented. Scatter decomposition divides an irregular computational domain into a large number of equal sized pieces, and distributes them modularly among processors. A probabilistic model of workload in one dimension is used to formally explain why, and when scatter decomposition works. The first result is that if correlation in workload is a convex function of distance, then scattering a more finely decomposed domain yields a lower average processor workload variance. The second result shows that if the workload process is stationary Gaussian and the correlation function decreases linearly in distance until becoming zero and then remains zero, scattering a more finely decomposed domain yields a lower expected maximum processor workload. Finally it is shown that if the correlation function decreases linearly across the entire domain, then among all mappings that assign an equal number of domain pieces to each processor, scatter decomposition minimizes the average processor workload variance. The dependence of these results on the assumption of decreasing correlation is illustrated with situations where a coarser granularity actually achieves better load balance.

Nicol, David M.↗

Deep learning for time series forecasting: a survey of recent advances

Time series forecasting plays a critical role in numerous real-world applications, such as finance, healthcare, transportation, and scientific computing. In recent years, deep learning has become a powerful tool for modeling complex temporal patterns and improving forecasting accuracy. This survey provides an overview of recent deep learning approaches for time series forecasting, involving various architectures including RNNs, CNNs, GNNs, transformers, large language models, MLP-based models, and diffusion models. We first identify key challenges in the field, such as temporal dependency, efficiency, and cross-variable dependency, which drive the development of forecasting techniques. Then, the general advantages and limitations of each architecture are discussed to contextualize their adaptation in time series forecasting. Furthermore, we highlight promising design trends like multi-scale modeling, decomposition, and frequency-domain techniques, which are shaping the future of the field. This paper serves as a compact reference for researchers and practitioners seeking to understand the current landscape and future trajectory of deep learning in time series forecasting.

97 MATHEMATICS AND COMPUTING↗

Parallel decomposition methods for the solution of electromagnetic scattering problems

This paper contains a overview of the methods used in decomposing solutions to scattering problems onto coarse-grained parallel processors. Initially, a short summary of relevant computer architecture is presented as background to the subsequent discussion. After the introduction of a programming model for problem decomposition, specific decompositions of finite difference time domain, finite element, and integral equation solutions to Maxwell's equations are presented. The paper concludes with an outline of possible software-assisted decomposition methods and a summary.

Cwik, Tom↗

Rank-Limiting Strategies for Optimizing Tensor-Train Finite-Difference Time-Domain Simulations

We introduce rank-limiting strategies to optimize tensor-train decompositions for three-dimensional finite-difference time-domain simulations using the relationship between the tensors and their specific dimensionality. These include the use of hard caps on the inner ranks of the tensor train decomposition and the use of a group rounding algorithm taking into account all field components simultaneously. Here, several numerical examples are considered to verify the efficacy of the proposed optimization strategies.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Implementation of a Message Passing Interface into a Cloud-Resolving Model for Massively Parallel Computing

The capability for massively parallel programming (MPP) using a message passing interface (MPI) has been implemented into a three-dimensional version of the Goddard Cumulus Ensemble (GCE) model. The design for the MPP with MPI uses the concept of maintaining similar code structure between the whole domain as well as the portions after decomposition. Hence the model follows the same integration for single and multiple tasks (CPUs). Also, it provides for minimal changes to the original code, so it is easily modified and/or managed by the model developers and users who have little knowledge of MPP. The entire model domain could be sliced into one- or two-dimensional decomposition with a halo regime, which is overlaid on partial domains. The halo regime requires that no data be fetched across tasks during the computational stage, but it must be updated before the next computational stage through data exchange via MPI. For reproducible purposes, transposing data among tasks is required for spectral transform (Fast Fourier Transform, FFT), which is used in the anelastic version of the model for solving the pressure equation. The performance of the MPI-implemented codes (i.e., the compressible and anelastic versions) was tested on three different computing platforms. The major results are: 1) both versions have speedups of about 99% up to 256 tasks but not for 512 tasks; 2) the anelastic version has better speedup and efficiency because it requires more computations than that of the compressible version; 3) equal or approximately-equal numbers of slices between the x- and y- directions provide the fastest integration due to fewer data exchanges; and 4) one-dimensional slices in the x-direction result in the slowest integration due to the need for more memory relocation for computation.

Juang, Hann-Ming Henry↗

Multiscale Characterization of Electrode-Induced Degradation in Perovskite Solar Cells

The stability of metal-halide-perovskite (MHP) solar cells must be understood and improved for the commercial viability of MHP technologies. Here, we apply multiscale characterization methods to study degradation modes, specifically electrode corrosion, for p-i-n MHP partial device stacks and full devices that are stored in the dark under an inert atmosphere. Our multiscale characterization approaches include full-device electro-optical performance using current-voltage (JV) curves and spatial imaging with electroluminescence (EL) and photoluminescence (PL). We further correlate interface properties using cross-sectional Kelvin probe force microscopy, which maps the nanoscale electric field properties, and electron microscopy, which demonstrates structural and chemical features. Devices stored as a full device stack degrade primarily by metal (Ag) electrode diffusion into the absorber, with formation of AgI byproducts and Ag accumulation near the indium tin oxide (ITO) contact. This causes decomposition of the perovskite absorber domains, loss of the potential drop at the electron transport layer (ETL)/perovskite interface near the metal contact, and increased equivalent resistance at the perovskite/hole transport layer (HTL) interface near the ITO contact. The devices stored without metal show a different degradation pathway dominated by corrosion of the ITO, creating voids at the ITO electrode surface with diffusion of In and Sn into the absorber. We conclude that metal electrode-induced degradation is the most severe degradation pathway under dark storage, but that ITO corrosion and absorber instability must also be mitigated. We further demonstrate mitigation of these degradation pathways by changes to the device stack, including a SnO x blocking layer at the ETL side and replacing ITO with FTO at the HTL side. These results provide a useful demonstration of specific dark degradation pathways at each electrode interface, as well as a unique multiscale example that links degradation of chemical, structural, and electrical interface properties to the full-device electro-optical characteristics.

14 SOLAR ENERGY↗

Goated: goal-oriented tensor decompositions in python

SAND2026-20464O Goated performs goal-oriented tensor decompositions in Python, enabling efficient compression of multi-dimensional simulation data. It extends common tensor decomposition methods by incorporating domain-specific knowledge, such as conservation laws in physics, through a penalty term in the optimization process. This approach improves data compression and modeling accuracy across various applications, including physics simulations, by using specialized algorithms and structure-aware subroutines to accelerate solver performance. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy's National Nuclear Security Administration under contract DE-NA0003525.

SciDAC↗

An exploration of function analysis and function allocation in the commercial flight domain

The applicability is explored of functional analysis methods to support cockpit design. Specifically, alternative techniques are studied for ensuring an effective division of responsibility between the flight crew and automation. A functional decomposition is performed of the commercial flight domain to provide the information necessary to support allocation decisions and demonstrate methodology for allocating functions to flight crew or to automation. The function analysis employed 'bottom up' and 'top down' analyses and demonstrated the comparability of identified functions, using the 'lift off' segment of the 'take off' phase as a test case. The normal flight mission and selected contingencies were addressed. Two alternative methods for using the functional description in the allocation of functions between man and machine were investigated. The two methods were compared in order to ascertain their relative strengths and weaknesses. Finally, conclusions were drawn regarding the practical utility of function analysis methods.

Mcguire, James C.↗

An analysis of scatter decomposition

A formal analysis of a mapping method known as scatter decomposition (SD) is presented. SD divides an irregular domain into many equal-size pieces and distributes them modularly among processors. It is shown that, if a correlation in workload is a convex function of distance, then scattering a more finely decomposed domain yields a lower average processor workload variance; if the workload process is stationary Gaussian and the correlation function decreases linearly in distance to zero and then remains zero, scattering a more finely decomposed domain yields a lower expected maximum processor workload. Finally, if the correlation function decreases linearly across the entire domain, then (among all mappings that assign an equal number of domain pieces to each processor) SD minimizes the average processor workload variance. The dependence of these results on the assumption of decreasing correlation is illustrated with cases where a coarser granularity actually achieves better load balance.

Nicol, David M.↗

Spectral Proper Orthogonal Decomposition of uPSP Measurements in Recent NASA Ames Wind Tunnel Test

This paper discusses Spectral Proper Orthogonal Decomposition (SPOD) of the Unsteady Pressure-Sensitive Paint (uPSP) measurements in recent NASA Ames wind tunnel test. The uPSP measurements were collected using Innovative Scientific Solutions, Inc. (ISSI) porous, fast-response pressure-sensitive paint, 40 ISSI four-inch air-cooled Light-Emitting Diodes, and 8 Phantom v2512 high-speed cameras at 10,000 frames per second in the uPSP Launch Vehicle Demonstration Test (LVDT) of the Space Launch System (SLS) vehicle in the 11-by 11-foot transonic test section of the Unitary Plan Wind Tunnel at NASA Ames Research Center in April 2024. SPOD is derived from a space-time proper orthogonal decomposition problem for statistically stationary flows. SPOD modes are determined in the frequency domain. Each SPOD mode oscillates at a single frequency. SPOD can be viewed as an extension of the Discrete Fourier Transform composition and the Dynamic Mode Decomposition. In this paper, the outputs of SPOD of the uPSP measurements in the uPSP LVDT are presented and the effectiveness of SPOD in the identification, diagnosis and analysis of the aerodynamic and aeroacoustic phenomena is demonstrated. The unsteady and dynamic property of the pressure field on the surface of the SLS Block 1B crew vehicle is presented with the visualization of the SPOD modes of the uPSP measurements in the tests of a Mach sweep run of the uPSP LVDT. The SPOD outputs were generated with the execution in parallel of a code in Python, with the library of Message Passing Interface for parallel processing, on the NASA Pleiades supercomputer. The work described in this paper is a part of NASA’s development of a new state-of-the-art uPSP capability in production wind tunnels. Funding was provided by the NASA Aerosciences Evaluation and Test Capabilities Portfolio Office.

Aeroacoustics↗

Enabling Grid-Forming Control with Fault Ride-Through in Unbalanced Distribution Networks

Distribution networks are often unbalanced, causing oscillatory responses in inverter control designed for balanced conditions. Here, to address this problem, this paper proposes a novel time-domain transformation appropriate for inverter control and enables the decomposition of three-phase unbalanced signals into constant positive and negative components. Relations useful for calculating unbalanced active and reactive power are derived from first principle, providing insight into vector products of unbalanced three-phase signals. Furthermore, a grid-forming control effective under unbalanced conditions is developed, which delivers superior performance while meeting UNIFI1 specifications for grid-forming control under unbalanced conditions. specifications applicable to category 4 inverter-based resource, like setting and regulating frequency/voltage, providing voltage support, sharing active power, injecting negative sequence current, and riding through faults. A current limiter is proposed for safe fault ride-through and integrates with the grid-forming control featuring frequency/voltage droop controllers and current and voltage control loops. The transformation of interconnected inverters is formulated and stability of the proposed control analyzed to support robust parameter selections. The effectiveness of the proposed transformation and grid-forming control is demonstrated through analytical results and real-time simulation of a IEEE 123 distribution network on the Real-Time Digital Simulator. Comparison with existing methods shows that the proposed strategy satisfies the UNIFI specifications with a much better performance.

24 - POWER TRANSMISSION AND DISTRIBUTION↗

Subband And Transform Compression Of Video Signals

Class of hierarchical subband coders developed primarily for compression of image data at video rates. Offers good performance with limited computational complexity and with flexibility inherent in subband decomposition. Particular subband decomposition chosen for coders appears to hide large quantitative errors effectively, largely because decomposition occurs along two-dimensional spatial-frequency-domain boundaries resembling spatial-frequency-domain curves of constant sensitivity of human visual system. Curves found approximately diamond-shaped: thus, low-pass filtering for reduction of data ideally involves nonrectangular passbands.

Sauer, Ken↗

The effect of inlet swirl on the dynamics of long annular seals in centrifugal pumps

This paper describes additional results from a continuing research program which aims to identify the dynamics of long annular seals in centrifugal pumps. A seal test rig designed at Heriot-Watt University and commissioned at Weir Pumps Research Laboratory in Alloa permits the identification of mass, stiffness, and damping coefficients using a least-squares technique based on the singular value decomposition method. The analysis is carried out in the time domain using a multi-fiequency forcing function. The experimental method relies on the forced excitation of a flexibly supported stator by two hydraulic shakers. Running through the stator embodying two symmetrical balance drum seals is a rigid rotor supported in rolling element bearings. The only physical connection between shaft and stator is the pair of annular gaps filled with pressurized water discharged axially. The experimental coefficients obtained from the tests are compared with theoretical values.

Ismail, M.↗

Octree based automatic meshing from CSG models

Finite element meshes derived automatically from solid models through recursive spatial subdivision schemes (octrees) can be made to inherit the hierarchical structure and the spatial addressability intrinsic to the underlying grid. These two properties, together with the geometric regularity that can also be built into the mesh, make octree based meshes ideally suited for efficient analysis and self-adaptive remeshing and reanalysis. The element decomposition of the octal cells that intersect the boundary of the domain is emphasized. The problem, central to octree based meshing, is solved by combining template mapping and element extraction into a procedure that utilizes both constructive solid geometry and boundary respresentation techniques. Boundary cells that are not intersected by the edge of the domain boundary are easily mapped to predefined element topology. Cells containing edges (and vertices) are first transformed into a planar polyhedron and then triangulated via element extractors. The modeling environments required for the derivation of planar polyhedra and for element extraction are analyzed.

Perucchio, Renato↗