Search NASASearch

SEARCH · Search NASA

Results for “Basis sets”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Physics-Driven Construction of Compact Primitive Gaussian Density Fitting Basis Sets

We present a model-assisted density fitting (MADF) basis set generator, an algorithm for generating primitive atomic Gaussian density fitting (DF) basis sets (DFBSs) from a contracted Gaussian orbital basis set (OBS). The MADF algorithm produces DFBSs suitable for accurate robust DF approximation of 2-particle interactions in mean-field and correlated electronic structures. The algorithm is designed to (a) saturate the OBS product space by a large regularized set of primitive solid-harmonic Gaussian shells with nonuniform distribution of exponents, followed by (b) pruning of the shells according to their contributions to the 2- body energy of a correlated atomic ensemble. Building the DFBS generator model almost exclusively on mathematical and physical principles allows one to limit the number of parameters that control the density fitting error to three, with a single set of parameters sufficient for computations with all basis cardinal numbers, with and without correlation of core electrons, with and without scalar and spin-dependent relativistic effects, spanning almost all of the Periodic Table. Performance assessment included basis sets up to quadruple-ζ quality from several major basis set families, using molecules composed of main-group, d-block, and f-block elements. The resulting DF errors in Hartree−Fock and second-order MP2 energies (with relativistic all-electron treatments, when appropriate) were on the order of 20 and 10 μE h per electron, respectively.

Approximation

Enhancing the accuracy of XPS calculations: Exploring hybrid basis set schemes for CVS-EOMIP-CCSD calculations

Reliable computational methodologies and basis sets for modeling x-ray spectra are essential for extracting and interpreting electronic and structural information from experimental x-ray spectra. In particular, the trade-off between numerical accuracy and computational cost due to the size of the basis set is a major challenge, since molecular orbitals undergo extreme relaxation in the core-hole state. To gain clarity on the changes in electronic structure induced by the formation of a core-hole, the use of sufficiently flexible basis for expanding the orbitals, particularly for the core region, has been shown to be essential. This work focuses on the refinement of core-hole ionized state calculations using the equation-of-motion coupled cluster family of methods through an extensive analysis on the effectiveness of “hybrid” and mixed basis sets. In this investigation, we utilize the CVS-EOMIP-CCSD method in combination and construct hybrid basis sets piecewise from readily available Dunning’s correlation consistent basis sets in order to calculate x-ray ionization energies (IEs) for a set of small gas phase molecules. Our results provide insights into the impact of basis sets on the CVS-EOMIP-CCSD calculations of K-edge IEs of first-row p-block elements. Furthermore, these insights enable us to understand more about the basis set dependence of the core IEs computed and allow us to establish a protocol for deriving reliable and cost-effective theoretical estimates for computing IEs of small molecules containing such elements.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Multiresolution Quantum Chemistry: Nonlinear Response Properties at the Basis Set Limit

We benchmark the accuracy of Dunning correlation-consistent Gaussian basis sets for computing frequencydependent second-order hyperpolarizabilities relevant to second-harmonic generation (SHG), using multiresolution analysis (MRA) as a reference. Basis set errors are analyzed using a unit-sphere representation of the effective hyperpolarizability vector, enabling direct assessment of directional error structure. We introduce a relative RMS total error metric that integrates directional deviations over the unit sphere and complement it with signed projection errors that distinguish over- and underestimation. Unsupervised clustering based on these signed directional metrics reveals four distinct convergence behaviors across a set of 68 molecules. Unitsphere visualizations of representative systems show that basis set errors are often highly anisotropic and localized along specific bond directions, even when global error measures appear small. Doubly augmented basis sets consistently outperform singly augmented ones, and core-polarization functions are required for uniform convergence in second-row systems. Overall, this work demonstrates that directional analysis combined with clustering provides a robust framework for understanding basis set convergence in nonlinear optical response properties.

Basis sets

Reducing the Cost of CCSD Basis Set Extrapolation in Ab Initio Computational Thermochemistry

Here, a series of approximations to CCSD contributions in computational model chemistries is presented in the context of kcal mol –1 , kJ mol –1 , and 20 cm –1 theoretical predictions of total atomization energies, benchmarked within the HEAT+CH 4 test suite. A specific set of circumstances where MP2, without empirical scaling, may be used as an effective intermediate in the first two of these accuracy ranges was determined. However, SDQ-MP4, a method long used in pursuit of kcal mol –1 accuracy but relatively unstudied in the subchemical accuracy community, offers significant improvement over the quality of MP2 as a basis-set intermediate at significantly reduced cost compared to CCSD. Given this, we argue for SDQ-MP4 as the de facto CCSD basis-set intermediate in sub-chemical accuracy calculations when CCSD in a desired basis set becomes unaffordable. We additionally report on a “CBS-like” scheme, where MP2 and SDQ-MP4 are used in conjunction to create a “cheap” three-part approximation of large CCSD basis set limits. The data for the CCSD approximation schemes are organized in such a way that model chemistry developers can locate an analog of their current approach for the CCSD basis set limit and explore alternative intermediates that either decrease computational cost or increase computational accuracy. We also show, for a handful of molecules, that SDQ-MP4 shows promise as an effective basis-set intermediate for harmonic and fundamental frequency computations, allowing for zero-point corrections of nearly CCSD(T)/ANO1 quality using simple composite methods that only require CCSD(T)/ANO0.

Thorpe, James H. [Argonne National Laboratory (ANL

Many-Body Basis Set Amelioration Method for Incremental Full Configuration Interaction

Incremental full configuration interaction (iFCI) is a polynomial-cost electronic structure method that systematically approaches the FCI limit by employing the method of increments to solve the Schrödinger equation through a many-body expansion. This article introduces the many-body basis set amelioration (MBBSA) method, which is designed to allow iFCI to be applicable to larger atomic orbital basis sets. MBBSA uses a series of inexpensive iFCI calculations to approximate the correlation energy that would be found using a more expensive, highly accurate iFCI calculation. Here, when compared to standard iFCI computations on smaller molecules in triple-zeta and larger basis sets, MBBSA provides approximations to the total and relative energies within chemical accuracy. MBBSA exhibits a reduced cost of between 60-92% when compared to standard iFCI calculations, with larger systems experiencing the largest benefit. Tests of MBBSA on two reactions that involve highly correlated systems, the automerization of cyclobutadiene and a Criegee intermediate reaction, show that MBBSA has practical utility for studying realistic chemistries.

Basis sets

Perturbative second-order optical susceptibility of bulk materials: a symmetry-enforced return to non-orthogonal localized basis sets

The second-order optical susceptibility of semiconductors $\chi^{(2)}_{ijk}(-2\omega;\omega,\omega)$ finds application in metrology, spectroscopy, telecommunications, material characterization, and quantum information. Pioneering calculations of $\chi^{(2)}_{ijk}(-2\omega;\omega,\omega)$ utilized non-orthogonal Gaussian orbitals centered at atoms. That formulation transitioned into plane-wave-based algorithms as time went by. As of late, nevertheless, multiple tools for calculating optical susceptibilities have recast the problem using Wannier (i.e. localized) orbitals, making a comeback onto frameworks based on localized basis sets. Here, in this work, we present an approach for calculating $\chi^{(2)}_{ijk}(-2\omega;\omega,\omega)$ reliant on numerical pseudo-atomic orbitals (PAOs) within perturbation theory in the velocity gauge. Its salient feature is a calculation of ‘Slater–Koster-like’ two-center integrals of the momentum operator in between PAOs identified by symmetry. The approach was successfully tested on paradigmatic cubic silicon carbide (3C-SiC) and gallium arsenide, for which linear responses are contributed as well.

Huamán, Angiolo [Univ. of Arkansas, Fayetteville,

Taming the virtual space for incremental full configuration interaction

Incremental full configuration interaction (iFCI) closely approximates the FCI limit with polynomial cost through a many-body expansion of the correlation energy, providing highly accurate total energies within a given basis set. To extend iFCI beyond previous basis set limitations, this work introduces a novel natural orbital (NO) screening approach, incremental NO full configuration interaction (iNO-FCI). By consideration of the importance of virtual orbital selection in the convergence of iFCI, iNO-FCI maximizes the consistency between orbitals selected for each correlated body. iNO-FCI employs a principle of cancellation of errors and ensures that the same set of virtual NOs is used for interdependent terms. Here, this strategy significantly reduces computational cost without compromising precision. Computational savings of up to 95% are demonstrated, allowing access to larger basis sets that were previously computationally prohibitive. iNO-FCI is herein introduced and benchmarked for several difficult test cases involving double-bond dissociation, biradical systems, conjugated π systems, and the spin gap of a Cu-based transition metal complex.

Correlation energy

Hydrogen Bond Benchmark: Focal‐Point Analysis and Assessment of DFT Functionals

We performed a hierarchical, convergent ab initio benchmark study and systematically analyzed the performance of density functional approximations for describing hydrogen bonds in small neutral, cationic, and anionic complexes, as well as in larger systems involving amide, urea, deltamide, and squaramide moieties. Focal point analyses (FPA), extrapolating to the ab initio limit, were carried out using correlated wave function methods up to CCSDT(Q) for the small complexes and CCSD(T) for the larger systems, together with correlation-consistent Gaussian basis sets up to the complete basis set limit. Optimized geometries and vibrational frequencies were obtained at the CCSD(T) level. The resulting FPA hydrogen-bond energies converge within a few tenths of a kcal mol −1 . These reference data were used to evaluate 60 density functionals (including 12 dispersion-corrected), spanning the local-density approximation (LDA), generalized gradient approximations (GGAs), meta-GGAs, hybrids, meta-hybrids, double-hybrids, and range-separated hybrids. Overall, the meta-hybrid M06-2X provides the best performance for both hydrogen bond energies and geometries, while the dispersion-corrected GGAs BLYP-D3(BJ) and BLYP-D4 also yield accurate hydrogen-bond data and can serve as cost-effective options for studying large and complex systems.

coupled cluster theory

The Electronic Structure of Zirconium and Hafnium Monochalcogenides

High-level ab initio CCSD(T) and spin–orbit icMRCI+Q calculations were used to predict potential energy curves (PECs) for the lowest-lying states of ZrO, ZrS, HfO, and HfS. The prediction of the ground state is basis set dependent at the icMRCI+Q level for ZrO and ZrS due to the small singlet–triplet splitting between the lowest 1 Σ + and 3 Δ states. CCSD(T) with a spin orbit correction predicted the 1 Σ + ground state in agreement with experiment. New all-electron basis sets were developed for Hf to improve the results over those predicted by use of effective core potentials (ECPs) that subsume the 4f electrons into the definition of the core. The use of the new DK-4f basis sets rather than ECPs became more important for HfO and HfS where there is a lack of a good core–valence separation. icMRCI+Q, CCSD(T), and DFT calculations for the spectroscopic parameters of ZrO, ZrS, HfO, and HfS were benchmarked with available experimental data. Bond dissociation energies (BDEs) of these four systems were calculated at the Feller–Peterson–Dixon (FPD) level to be 762.1 (ZrO), 543.5 (ZrS), 803.8 (HfO), and 575.1 kJ/mol (HfS), in excellent agreement with experiment. The HfS BDE was remeasured using the R3PI method, providing an updated experimental measurement of D 0 (HfS) = 5.978 ± 0.002 eV = 576.8 ± 0.2 kJ/mol. This experimental value, combined with experimental measurements of the ionization energies of Hf and HfS, gives the cationic BDE of D 0 (Hf + -S) = 5.124 ± 0.002 eV = 494.4 ± 0.2 kJ/mol.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Slimmer Geminals For Accurate F12 Electronic Structure Models

The Slater-type F12 geminal length scales originally tuned for the second-order Mo̷ller-Plesset F12 method are too large for higher-order F12 methods formulated using the SP (diagonal fixed-coefficient spin-adapted) F12 ansatz. The new geminal parameters reported herein reduce the basis set incompleteness errors (BSIEs) of absolute coupled-cluster singles and doubles F12 correlation energies by a significant─and increase with the cardinal number of the basis─margin. The effect of geminal reoptimization is especially pronounced for the cc-pVXZ-F12 basis sets (specifically designed for use with F12 methods) relative to their conventional aug-cc-pVXZ counterparts. The BSIEs of relative energies are less affected, but substantial reductions can be obtained, especially for atomization energies and ionization potentials with the cc-pVXZ-F12 basis sets. The new geminal parameters are therefore recommended for all applications of high-order F12 methods, such as coupled-cluster F12 methods and transcorrelated F12 methods.

Powell, Samuel R. [Virginia Polytechnic Inst. and

An Algorithm for Atom-Centered Lossy Compression of the Atomic Orbital Basis in Density Functional Theory Calculations

Large atomic-orbital (AO) basis sets of at least triple and preferably quadruple-ζ (QZ) size are required to adequately converge Kohn–Sham density functional theory (DFT) calculations toward the complete basis set limit. However, incrementing the cardinal number by one nearly doubles the AO basis dimension, and the computational cost scales as the cube of the AO dimension, so this is very computationally demanding. Here, in this work, we develop and test a threshold-based natural atomic orbital (NAO) scheme in which ϵ-NAOs are obtained as eigenfunctions of atomic blocks of the density matrix in a one-center orthogonalized representation. This enables compression of the AO basis that is optimal for a given threshold, 10 –ϵ , by discarding NAOs with occupation numbers below that threshold. Extensive pilot test calculations using the Hartree–Fock functional and taking the converged density matrix as input suggest that a threshold of 10 –5 can yield a compression factor (ratio of AO to compressed ϵ-NAO dimension) between 2.5 and 4.5 for the QZ pc-3 basis. The errors in relative energies are typically less than 0.1 kcal/mol when the compressed basis is used instead of the uncompressed basis. Between 10 and 100 times smaller errors (i.e., usually less than 0.01 kcal/mol) can be obtained with a threshold 10 –7 , while the compression factor is typically between 2 and 2.5.

basis sets

Convergent Protocols for Computing Protein–Ligand Interaction Energies Using Fragment-Based Quantum Chemistry

Fragment-based quantum chemistry methods offer a way to sidestep the steep nonlinear scaling of electronic structure calculations so that large molecular systems can be investigated using high-level methods. Here, we use fragmentation to compute protein–ligand interaction energies in systems with several thousand atoms, using a new software platform for managing fragment-based calculations that implements a screened many-body expansion. Convergence tests using a minimal-basis semiempirical method (HF-3c) indicate that two-body calculations, with single-residue fragments and simple hydrogen caps, are sufficient to reproduce interaction energies obtained using conventional supramolecular electronic structure calculations, to within 1 kcal/mol at about 1% of the computational cost. We also demonstrate that the HF-3c results are illustrative of trends obtained with density functional theory in basis sets up to augmented quadruple-ζ quality. Strategic deployment of fragmentation facilitates the use of converged biomolecular model systems alongside high-quality electronic structure methods and basis sets, bringing ab initio quantum chemistry to systems of hitherto unimaginable size. This will be useful for generation of high-quality training data for machine learning applications.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Method-independent cusps for atomic orbitals in quantum Monte Carlo

Here, we present an approach for augmenting Gaussian atomic orbitals with correct nuclear cusps. Like the atomic orbital basis set itself and unlike previous cusp corrections, this approach is independent of the many-body method used to prepare wave functions for quantum Monte Carlo. Once the basis set and molecular geometry are specified, the cusp-corrected atomic orbitals are uniquely specified, regardless of which density functionals, quantum chemistry methods, or subsequent variational Monte Carlo optimizations are employed. We analyze the statistical improvement offered by these cusps in a number of molecules and find them to offer similar advantages as molecular-orbital-based approaches while remaining independent of the choice of many-body method.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Convergent Concordant Mode Approach for Molecular Vibrations: CMA-2

The concordant mode approach (CMA) is a promising new scheme for dramatically increasing the system size and level of theory achievable in quantum chemical computations of molecular vibrational frequencies. Here, we achieve advances in the CMA hierarchy by computations targeting CCSD(T)/cc-pVTZ (coupled cluster singles and doubles with perturbative triples using a correlation-consistent polarized-valence triple-ζ basis set) benchmarks within the G2 molecular test set, executing a statistical analysis for 1501 frequencies from 111 compounds and then separately solving the refractory case of pyridine. First, MP2/cc-pVTZ (second-order Møller–Plesset perturbation theory with the same basis set) proves to be an excellent and preferred choice for generating the underlying (Level B) normal modes of the CMA scheme. Utilizing this Level B within the CMA-0A method reproduces the 1501 benchmark frequencies with a mean absolute error (MAE) of only 0.11 cm –1 and an attendant standard deviation of 0.49 cm –1 . Second, a convergent CMA-2 method is constituted that allows efficient computation of higher level (Level A) frequencies to any reasonable accuracy threshold by using only Hartree–Fock (HF) and MP2 or density functional theory (DFT) data to generate ξ parameters, which select the sparse off-diagonal force field elements for explicit evaluation at Level A. When Level B = MP2/cc-pVTZ, a cutoff of ξ = 0.02 provides an average maximum absolute error per molecule of only 0.17 cm –1 by incurring merely a 33% increase in average cost over CMA-0A. This CMA-2 method also eradicates the 4 problematic CMA-0A outliers of pyridine with even less effort (ξ = 0.04, 22% increase). Finally, the newly developed CMA procedures are shown to be highly successful when applied to 1-(1H-pyrrol-3-yl)ethanol, a new test molecule with diverse types of vibration.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Prospects for detecting UF6 hydrolysis intermediates by mass spectrometry, resonance Raman, and NQR spectroscopy

Simulations are performed to consider several spectrometric and spectroscopic candidates for elucidating the mechanism of the hydrolysis of uranium hexafluoride (UF6). This study is among the first to benchmark the def-mTZVP basis sets for actinide-containing molecules, and it is shown to be a suitable basis set for surveying geometrical structures and vibrational spectra when used in conjunction with density functional theory. An experiment is proposed coupling mass spectrometric ion selection with vibrational spectroscopy, and supporting infrared and Raman spectral simulations demonstrate that ionization blue-shifts bands and can change their qualitative features. Ultraviolet resonance Raman is shown to have good prospects for discriminating U–O–U bridged intermediates, evidence for which was observed recently [L. E. McNamara et al., J. Phys. Chem. A 130, 775–786 (2026)]. As a more speculative approach, we also consider the prospect of addressing 233U or 235U nuclei by quadrupole resonance (NQR) spectroscopy for assigning early stage intermediate complexes. In doing so, we provide a first order-of-magnitude estimate for the collision-induced NQR signal for the UF6 dimer, which is of interest for the interpretation of a fast decoherence time observed in liquid-phase nuclear magnetic resonance. Having provided critical insights into the formation and stability of UF6 hydrolysis intermediates in previous studies, this computational spectroscopy survey is expected to help guide and expedite future laboratory campaigns.

Lutz, Jesse J. [Center for Computing Research, San

Multi-fidelity learning for interatomic potentials: low-level forces and high-level energies are all you need

The promise of machine learning interatomic potentials (MLIPs) has led to an abundance of public quantum mechanical (QM) training datasets. The quality of an MLIP is directly limited by the accuracy of the energies and atomic forces in the training dataset. Unfortunately, most of these datasets are computed with relatively low-accuracy QM methods, e.g. density functional theory with a moderate basis set. Due to the increased computational cost of more accurate QM methods, e.g. coupled-cluster theory with a complete basis set (CBS) extrapolation, most high-accuracy datasets are much smaller and often do not contain atomic forces. The lack of high-accuracy atomic forces is quite troubling, as training with force data greatly improves the stability and quality of the MLIP compared to training to energy alone. Because most datasets are computed with a unique level of theory, traditional single-fidelity (SF) learning is not capable of leveraging the vast amounts of published QM data. In this study, we apply multi-fidelity learning (MFL) to train an MLIP to multiple QM datasets of different levels of accuracy, i.e. levels of fidelity. Specifically, we perform three test cases to demonstrate that MFL with both low-level forces and high-level energies yields an extremely accurate MLIP—far more accurate than a SF MLIP trained solely to high-level energies and almost as accurate as a SF MLIP trained directly to high-level energies and forces. Therefore, MFL greatly alleviates the need for generating large and expensive datasets containing high-accuracy atomic forces and allows for more effective training to existing high-accuracy energy-only datasets. Indeed, low-accuracy atomic forces and high-accuracy energies are all that are needed to achieve a high-accuracy MLIP with MFL.

36 MATERIALS SCIENCE

Activation of methane by U + studied by guided ion beam tandem mass spectrometry and quantum chemistry

Reaction pathways of all products formed in the U + + CH 4 (CD 4 ) reaction were explored as a function of kinetic energy using guided ion beam tandem mass spectrometry and quantum chemical calculations. UH + , UC + , UCH + , UCH 2 + , and UCH 3 + (and their perdeuterated analogues) are formed in endothermic reactions. In both systems, the UCH 2 + (UCD 2 + ) dehydrogenated product was the dominant product in the low-energy region, whereas the UH + (UD + ) hydride product became predominant at high energies. The kinetic energy behavior of the various products is consistent with a common intermediate of H–U + –CH 3 (D–U + –CD 3 ). Here, the kinetic energy dependence of all product cross sections was modeled to obtain experimental bond dissociation energies at 0 K (in eV): D 0 (U + –H) = 2.42 ± 0.10, D 0 (U + –C) = 3.95 ± 0.12, D 0 (U + –CH) = 4.91 ± 0.09, D 0 (U + –CH 2 ) = 4.11 ± 0.04, and D 0 (U + –CH 3 ) = 2.41 ± 0.09. Quantum chemical calculations using the UCCSD(T) and UB3LYP approaches with the cc-pwCVXZ-PP basis set with MDF-60 pseudopotential for U + and the aug-cc-pCVXZ and aug-cc-pVXZ (X = T, Q) basis set for carbon and hydrogen, respectively, validate the experimental bond dissociation energies and outline the potential energy surface for all reactions observed. In addition, spin–orbit corrections of the bond energies for all products were calculated at a CASSCF-CASPT2-RASSI level.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH