Search NASASearch

SEARCH · Search NASA

Results for “text analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Natural Language Processing to Inform Agent-Based Modeling: With Application to Modeling Adoption of Medium-Duty Electric Vehicles

Agent-based socio-technical modeling of medium- and heavy-duty (MDHD) electric vehicle (EV) adoption has the potential to provide analysis, prediction, and gui. This paper describes new applications of text analysis developed through machine learning (ML) to build and understand relevant topics and their saliency in the published discourse on adoption of MDHD EVs. This work contributes to the state of the art in topic mining models by defining a new metric of topic ranking (START) that quantifies the importance of predefined topics within the corpus using weighted results for predefined topics from two topic modeling approaches: Latent Dirichlet Allocation (LDA) and BERTopic. The START metric is then demonstrated in practice to model how academia and industry view the EV adoption process based on the respective texts published by these groups. Results show that academic literature places more emphasis on categories of interests such as norms/attitudes and adopter knowledge, while trade journals tend to emphasize long-term cost more than academia. The two bodies of literature agree on the importance of policy and incentives in MDHD EV adoption. Together these results illustrate the potential to use ML-based text analysis to populate the characteristics of agent-based socio-technical models.

Electric vehicle adoption, fleet electrification,

Fast Active-Set Thresholding Method for Nonnegative Least Squares

Nonnegative Least Squares (NNLS) is a fundamental constrained optimization problem encountered in many applications such as image deblurring, signal processing, nonnegative matrix factorization, magnetic microscopy, and hyperspectral imaging. Active-set based methods are a common class of algorithms for solving NNLS which identify the optimal variable set of the NNLS solution. They do so by iteratively solving a series of unconstrained least squares problems, identifying which variables violate the nonnegativity constraints, and then swapping variables in/out of consideration until the optimal set of variables is found. Several variations improving upon this method exist in the literature. In this work, we propose an active-set swap heuristic which further improves upon existing active-set based methods for NNLS. Our optimizations are based upon adding multiple variables to the passive set within a threshold of the smallest gradient value and removing variables within a similar threshold of the closest boundary constraint. We leverage these optimizations to yield a Fast Active-Set Thresholding NNLS (FAST-NNLS) algorithm which significantly outperforms the existing state-of-the-art NNLS algorithms for a wide range of problems. Rigorous convergence guarantees are proven for the proposed method. We demonstrate the effectiveness of our proposed method on multiple synthetic datasets and two realworld text analysis applications. In doing so, we present the most comprehensive NNLS solver comparison in the literature to date.

Cobb, Benjamin [Georgia Institute of Technology]

Demystifying the Resilience of Large Language Models: An End-to-End Perspective

Deep neural networks are known to be resilient to random bit-wise faults in their parameters. However, this resilience has primarily been established through evaluations of classification models. The extent to which this claim holds for large-language models remains underexplored. In this work, we conduct an extensive measurement study on the impact of random bitwise faults in commercial-scale language models. We perform an in-depth analysis of the resulting generation outputs. We first expose that these language models are not truly resilient to random bit-flips. While aggregate metrics such as accuracy may suggest resilience, an in-depth inspection of the generated outputs shows significant degradation in text quality. Our analysis also shows that tasks requiring more complex reasoning suffer more from performance and quality degradation. Moreover, we extend our analysis to models with augmented reasoning capabilities, such as Chain-of-Thought or Mixture of Experts architectures, and characterize their failure scenarios under random bit-flips.

Sun, Yu

Precursor reaction pathway leading to BiFeO 3 formation: insights from text-mining and chemical reaction network analyses

BiFeO 3 (BFO) is a next-generation non-toxic multiferroic material with applications in sensors, memory devices, and spintronics, where its crystallinity and crystal structure directly influence its functional properties. Designing sol–gel syntheses that result in phase-pure BFO remains a challenge due to the complex interactions between metal complexes in the precursor solution. Here, we combine text-mined data and chemical reaction network (CRN) analysis to obtain novel insight into BFO sol–gel precursor chemistry. We perform text-mining analysis of 340 synthesis recipes with the emphasis on phase-pure BFO and identify trends in the use of precursor materials, including that nitrates are the preferred metal salts, 2-methoxyethanol (2 ME) is the dominant solvent, and adding citric acid as a chelating agent frequently leads to phase-pure BFO. Our CRN analysis reveals that the thermodynamically favored reaction mechanism between bismuth nitrate and 2ME interaction involves partial solvation followed by dimerization, contradicting assumptions in previous literature. We suggest that further oligomerization, facilitated by nitrite ion bridging, is critical for achieving the pure BFO phase.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Leveraging large language models to automate the identification of healthcare access barriers for veterans

Objective: To develop and evaluate an automated system for identifying healthcare barriers focusing on transportation issues in veterans’ clinical notes using large language models (LLMs) and to assess the impact of different prompting strategies on classification performance and explanation consistency. Methods: We developed a hybrid system combining pattern matching for templated notes with LLM analysis for free-text notes. Using 2000 manually annotated clinical notes, we compared four prompting strategies (dual-role short, dual-role long, analysis-first, analysis-only) across Mistral-7B and Llama-3.1 models. We evaluated classification performance using standard metrics and assessed explanation consistency through embedding similarity analysis. Results: The analysis-first strategy achieved superior performance, with Mistral-7B reaching an F1 score of 0.914, outperforming traditional machine learning approaches (GBM: 0.786, BERT: 0.811). LLMs demonstrated higher explanation consistency within models (mean cosine similarity 0.887–0.908) compared to cross-model similarities (0.767–0.872). Pattern matching successfully handled 6.7% of templated notes deterministically. Mistral-7B showed greater internal consistency but higher abstention rates compared to Llama-3.1. Conclusion: Requiring LLMs to analyze evidence before classification improves both accuracy and explanation consistency for identifying transportation barriers in clinical notes. This approach enables automated barrier detection at scale while providing clinically relevant explanations, supporting both population-level healthcare planning and individual patient care decisions.

Healthcare access barriers

AI-powered topic modeling: comparing LDA and BERTopic in analyzing opioid-related cardiovascular risks in women

Topic modeling is a crucial technique in natural language processing (NLP), enabling the extraction of latent themes from large text corpora. Traditional topic modeling, such as Latent Dirichlet Allocation (LDA), faces limitations in capturing the semantic relationships in the text document although it has been widely applied in text mining. BERTopic, created in 2022, leveraged advances in deep learning and can capture the contextual relationships between words. In this work, we integrated Artificial Intelligence (AI) modules to LDA and BERTopic and provided a comprehensive comparison on the analysis of prescription opioid-related cardiovascular risks in women. Opioid use can increase the risk of cardiovascular problems in women such as arrhythmia, hypotension etc. 1,837 abstracts were retrieved and downloaded from PubMed as of April 2024 using three Medical Subject Headings (MeSH) words: “opioid,” “cardiovascular,” and “women.” Machine Learning of Language Toolkit (MALLET) was employed for the implementation of LDA. BioBERT was used for document embedding in BERTopic. Eighteen was selected as the optimal topic number for MALLET and 23 for BERTopic. ChatGPT-4-Turbo was integrated to interpret and compare the results. The short descriptions created by ChatGPT for each topic from LDA and BERTopic were highly correlated, and the performance accuracies of LDA and BERTopic were similar as determined by expert manual reviews of the abstracts grouped by their predominant topics. The results of the t-SNE (t-distributed Stochastic Neighbor Embedding) plots showed that the clusters created from BERTopic were more compact and well-separated, representing improved coherence and distinctiveness between the topics. Our findings indicated that AI algorithms could augment both traditional and contemporary topic modeling techniques. In addition, BERTopic has the connection port for ChatGPT-4-Turbo or other large language models in its algorithm for automatic interpretation, while with LDA interpretation must be manually, and needs special procedures for data pre-processing and stop words exclusion. Therefore, while LDA remains valuable for large-scale text analysis with resource constraints, AI-assisted BERTopic offers significant advantages in providing the enhanced interpretability and the improved semantic coherence for extracting valuable insights from textual data.

Research & Experimental Medicine

Postdoctoral insights on mentoring excellence: a framework for best practices at Sandia National labs

The Sandia National Laboratories Strategic Plan FY24-FY27, updated for FY25, outlines Sandia’s two Big Labs-wide Goals, Accelerate Innovation and Lead in Modern Engineering. The goal of Accelerate Innovation is that “by FY27, Sandia will be a leader in scientific, engineering and operational innovation and an employer of choice for highly innovative and creative talent.” Sandia’s postdocs are leaders in innovation, well-versed in emerging techniques and cutting-edge methods, and capable of acting as a highly agile technical force across domains at the lab. As a federally funded research and development center (FFRDC), Sandia National Laboratories attracts top doctoral talent by offering a unique opportunity for postdoctoral researchers to develop at the crossroads of government, academia, and industry, working in multi-disciplinary teams and performing cutting-edge, mission-specific research that responds to immediate needs of national interest. However, this creates unique opportunities and demands of both postdoctoral appointees and the Sandia staff who act as their mentors, making mentorship key to attract talent. Since 2007 the Sandia Postdoctoral Development (SPD) Board, originally Postdoc To Professional (PD2P), a networking group at Sandia composed of a voluntary board of current postdocs and two staff liaisons, has advocated for postdoctoral development within Sandia National Labs. In this white paper, SPD board members and the Sandia Postdoctoral Development Office, organized in 2019, have come together to develop a comprehensive overview of the postdoctoral mentoring landscape at Sandia National Labs as we currently know it. By scouring various forms of data from efforts since 2018, we’ve compiled a community-derived perspective on what makes postdoctoral mentorship at Sandia unique. First, we analyze working sessions held between mentors and mentees to develop a comprehensive map of who is involved in postdoctoral mentorship at the lab and how the responsibilities are divided amongst mentors and mentees. We then combine multiple forms of data, including exit surveys, annual surveys, and community workshops, to identify the specific challenges that mentors and mentees encounter at the national lab. Finally, we use text mining and sentiment analysis to analyze mentoring award data to develop an idea of what postdocs are self-identifying as excellent mentorship within the lab. It is our goal that this white paper act as an ongoing resource to the postdoc and postdoc mentoring communities and provide a firm foundation for further conversations on the future of postdoctoral mentorship at Sandia National Labs.

99 GENERAL AND MISCELLANEOUS

Transitory sensitivity in automatic chemical kinetic mechanism analysis

Abstract Detailed chemical kinetic mechanisms are necessary for resolving many important chemical processes. As the chemistry of smaller molecules has become better grounded and quantum chemistry calculations have become cheaper, kineticists have become interested in constructing progressively larger kinetic mechanisms to model increasingly complex chemical processes. These large kinetic mechanisms prove incredibly difficult to refine and time‐consuming to interpret. Traditional sensitivity analysis on a large mechanism can range from inconvenient to practically impossible without special techniques to reduce the computational cost. We first present a new time‐local sensitivity analysis we term transitory sensitivity analysis. Transitory sensitivity analysis is demonstrated in an example to accurately identify traditionally sensitive reactions at an 18,000x speed up over traditional sensitivities. By fusing transitory sensitivity analysis with more traditional time‐local branching, pathway, and cluster analyses, we develop an algorithm for efficient automatic mechanism analysis. This automatic mechanism analysis at a time point is able to identify the reactions a target is most sensitive to using transitory sensitivity analysis and then propose hypotheses why the reaction might be sensitive using branching, pathway, and cluster analyses. We implement these algorithms within the reaction mechanism simulator (RMS) package, which enables us to report the automatic mechanism analysis results in highly readable text formats and in molecular flux diagrams.

Johnson, Matthew S.

PhysBERT: A text embedding model for physics scientific literature

The specialized language and complex concepts in physics pose significant challenges for information extraction through Natural Language Processing (NLP). Central to effective NLP applications is the text embedding model, which converts text into dense vector representations for efficient information retrieval and semantic analysis. In this work, we introduce PhysBERT, the first physics-specific text embedding model. Pre-trained on a curated corpus of 1.2 × 106 arXiv physics papers and fine-tuned with supervised data, PhysBERT outperforms leading general-purpose models on physics-specific tasks, including the effectiveness in fine-tuning for specific physics subdomains.

Hellert, Thorsten (ORCID:0000000227970926)

Observations of the Marine Atmospheric Boundary Layer’s Response to a Solar Eclipse

The atmospheric response to the solar eclipse of 8 April 2024 in North America is investigated with a specific focus on the marine atmospheric boundary layer (MABL). We leverage measurements collected during the Third Wind Forecast Improvement Project (WFIP3), including Doppler lidars, sonic anemometers, and thermodynamic profiler data to investigate the atmospheric response across sites that experienced partial eclipse conditions with nearly 90% obscuration. Using these measurements, we examine eclipse-induced changes in key meteorological parameters, such as temperature, wind speed, and turbulent fluxes. Most previous eclipse studies have been conducted over land, whereas this study provides new observations for both coastal and marine environments, offering additional insight into eclipse-driven variability in the MABL. The findings confirm a notable decrease in downwelling shortwave radiation during the eclipse, which results in rapid cooling of surface air. The temperature reduction ranges from $1.2^\circ \text {C}$ to $1.4^\circ \text {C}$ in coastal regions and from $0.3^\circ \text {C}$ to $0.5^\circ \text {C}$ over the ocean. This analysis suggests that the MABL’s higher thermal inertia compared to coastal regions moderates the temperature decrease during the eclipse. Wind speed exhibits a more complex behavior, as it is influenced by both the MABL and preexisting synoptic conditions. Although a reduction in wind speed is observable up to approximately 140 m above ground level (AGL) at more inland sites, at other locations closer to the coast, this reduction is constrained to the lowest 100 m AGL. Turbulence parameters retrieved from sonic anemometers, such as turbulence kinetic energy, turbulent heat flux, and friction velocity, decrease during the eclipse at coastal sites, accompanied by a brief transition of atmospheric stability from unstable to neutral or weakly stable conditions. For the open-ocean sites, the variability in turbulence statistics and atmospheric stability is minimal during the occurrence of the eclipse.

16 TIDAL AND WAVE POWER

Simulation of Divertor Performance in ST40 Under Dynamic Double-Null Plasmas

A power fraction model was implemented for the simultaneous prediction of 3-D surface temperature evolution at all four divertor targets in near-double-null (DN) tokamak configurations, which is especially important for compact high-field devices that may not have the ability to dissipate large amounts of power on the high-field side. Evaluating the power-sharing between the four divertor strike points in a disconnected DN configuration is important for understanding the overall power balance, as well as for optimizing the power exhaust performance and prolonging the survivability of the plasma facing components (PFCs). This power-sharing is typically evaluated in terms of the separation between the primary and secondary separatrices at the outboard midplane, $\textit {dR}_{\text {sep}}$. The Heat flux Engineering Analysis Toolkit (HEAT) is coupled with Brunner’s power fraction model to simulate the deposited heat flux and resultant temperature change on 3-D divertor targets in a dynamic DN (DDN) pulse operation in ST40, a high-field spherical tokamak. The simulation results showed that with DDN operation, the operation time has significantly increased compared with single-null geometry configurations.

ST40

Determining Optimal Magnetometer Configuration on MAGIS-100

Long-baseline atom interferometers such as the Matter-wave Atomic Gradiometer Interferometric Sensor (MAGIS-100) require stringent control and continuous characterization of background magnetic fields and spatial gradients to prevent systemic phase shifts that mimic ultralight dark matter or gravitational wave signatures. Because direct sensor placement within the ultra-high vacuum beam pipe is infeasible, in-situ magnetic field monitoring relies on external sensor arrays situated in the surrounding annular region. This work demonstrates a field reconstruction framework for a 5.3-meter MAGIS-100 modular section using finite-element Opera simulations. Transverse magnetic fields are expanded using a cylindrical multipole framework as informed by Fermilab’s Muon g-2 experiment, with magnetometer array configurations optimized via Fisher information matrix D-optimality. Inverting external sensor readings through a Gauss-Newton scheme recovers interior tube fields across distinct axial positions. In the discontinuity-averse uniform region (slice pair P4), the model achieves sub-noise-floor performance with a cross-validated root-mean-square error (RMSE) of $6.7227 \times 10^{-4}\text{ A/m}$ ($0.845\times$ sensor noise floor) and an interior field coefficient of variation of $1.71\%$. An elbow criterion in the Fisher bounds establishes $n_{\text{max}} = 2$ as the optimal multipole truncation order to prevent noise amplification from over-parameterization, with $n_{\text{max}} = 3$ (sextupole) order chosen for analysis to demonstrate further complexity and cross-pair comparison. Furthermore, analytical differentiation of the fitted multipole coefficients yields dense spatial maps of the transverse Jacobian gradient matrix $\nabla \mathbf{H}$ along with propagated $1\sigma$ uncertainty bounds across the beam region ($r \le 2.75\text{ in}$). This operational framework confirms that external magnetometer arrays can reliably monitor magnetic field uniformity and spatial gradients along the 100-meter flight path given appropriate sampling for any complexity order.

Appleby, Darwin [William Rainey Harper Coll.] (ORC

Determining Optimal Magnetometer Configuration on MAGIS-100

Long-baseline atom interferometers such as the Matter-wave Atomic Gradiometer Interferometric Sensor (MAGIS-100) require stringent control and continuous characterization of background magnetic fields and spatial gradients to prevent systemic phase shifts that mimic ultralight dark matter or gravitational wave signatures. Because direct sensor placement within the ultra-high vacuum beam pipe is infeasible, in-situ magnetic field monitoring relies on external sensor arrays situated in the surrounding annular region. This work demonstrates a field reconstruction framework for a 5.3-meter MAGIS-100 modular section using finite-element Opera simulations. Transverse magnetic fields are expanded using a cylindrical multipole framework as informed by Fermilab’s Muon g-2 experiment, with magnetometer array configurations optimized via Fisher information matrix D-optimality. Inverting external sensor readings through a Gauss-Newton scheme recovers interior tube fields across distinct axial positions. In the discontinuity-averse uniform region (slice pair P4), the model achieves sub-noise-floor performance with a cross-validated root-mean-square error (RMSE) of $6.7227 \times 10^{-4}\text{ A/m}$ ($0.845\times$ sensor noise floor) and an interior field coefficient of variation of $1.71\%$. An elbow criterion in the Fisher bounds establishes $n_{\text{max}} = 2$ as the optimal multipole truncation order to prevent noise amplification from over-parameterization, with $n_{\text{max}} = 3$ (sextupole) order chosen for analysis to demonstrate further complexity and cross-pair comparison. Furthermore, analytical differentiation of the fitted multipole coefficients yields dense spatial maps of the transverse Jacobian gradient matrix $\nabla \mathbf{H}$ along with propagated $1\sigma$ uncertainty bounds across the beam region ($r \le 2.75\text{ in}$). This operational framework confirms that external magnetometer arrays can reliably monitor magnetic field uniformity and spatial gradients along the 100-meter flight path given appropriate sampling for any complexity order.

Appleby, Darwin [William Rainey Harper Coll.] (ORC

In-Situ Magnetic Field Reconstruction in the MAGIS-100 Experiment

Long-baseline atom interferometers such as the Matter-wave Atomic Gradiometer Interferometric Sensor (MAGIS-100) require stringent control and continuous characterization of background magnetic fields and spatial gradients to prevent systemic phase shifts that mimic ultralight dark matter or gravitational wave signatures. Because direct sensor placement within the ultra-high vacuum beam pipe is infeasible, in-situ magnetic field monitoring relies on external sensor arrays situated in the surrounding annular region. This work demonstrates a field reconstruction framework for a 5.3-meter MAGIS-100 modular section using finite-element Opera simulations. Transverse magnetic fields are expanded using a cylindrical multipole framework as informed by Fermilab’s Muon g-2 experiment, with magnetometer array configurations optimized via Fisher information matrix D-optimality. Inverting external sensor readings through a Gauss-Newton scheme recovers interior tube fields across distinct axial positions. In the discontinuity-averse uniform region (slice pair P4), the model achieves sub-noise-floor performance with a cross-validated root-mean-square error (RMSE) of $6.7227 \times 10^{-4}\text{ A/m}$ ($0.845\times$ sensor noise floor) and an interior field coefficient of variation of $1.71\%$. An elbow criterion in the Fisher bounds establishes $n_{\text{max}} = 2$ as the optimal multipole truncation order to prevent noise amplification from over-parameterization, with $n_{\text{max}} = 3$ (sextupole) order chosen for analysis to demonstrate further complexity and cross-pair comparison. Furthermore, analytical differentiation of the fitted multipole coefficients yields dense spatial maps of the transverse Jacobian gradient matrix $\nabla \mathbf{H}$ along with propagated $1\sigma$ uncertainty bounds across the beam region ($r \le 2.75\text{ in}$). This operational framework confirms that external magnetometer arrays can reliably monitor magnetic field uniformity and spatial gradients along the 100-meter flight path given appropriate sampling for any complexity order.

Appleby, Darwin [William Rainey Harper Coll.; Ferm

Cosmological constraints from the Planck cluster catalogue with DES shear profiles and Chandra observations

We present cosmological constraints from the Planck PSZ2 cosmological cluster sample, using weak-lensing shear profiles from Dark Energy Survey (DES) data and X-ray observations from the Chandra telescope for the mass calibration. We compute hydrostatic mass estimates for all clusters in the PSZ2 sample with a scaling relation between their Sunyaev-Zeldovich signal and X-ray derived hydrostatic mass, calibrated with the Chandra data. We introduce a method to correct these masses with a hydrostatic mass bias using shear profiles from wide-field galaxy surveys. We simultaneously fit the number counts of the PSZ2 sample and the mass calibration with the DES data, finding $Ω_\text{m}=0.312^{+0.018}_{-0.024}$, $σ_8=0.777\pm 0.024$, $S_8\equiv σ_8 \sqrt{Ω_\text{m} / 0.3}=0.791^{+0.023}_{-0.021}$, and $(1-b)=0.844^{+0.055}_{-0.062}$ for our baseline analysis when combined with BAO data. When considering a hydrostatic mass bias evolving with mass, we find $Ω_\text{m}=0.353^{+0.025}_{-0.031}$, $σ_8=0.751\pm 0.023$, and $S_8=0.814^{+0.019}_{-0.020}$. We verify the robustness of our results by exploring a variety of analysis settings, with a particular focus on the definition of the halo centre used for the extraction of shear profiles. We compare our results with a number of other analyses, in particular two recent analyses of cluster samples obtained from SPT and eROSITA data that share the same mass calibration data set. We find that our results are in overall agreement with most late-time probes, in very mild tension with CMB results (1.6$σ$), and in significant tension with results from eROSITA clusters (2.9$σ$). We confirm that our mass calibration is consistent with the eROSITA analysis by comparing masses for clusters present in both Planck and eROSITA samples, eliminating it as a potential cause of tension.

Aymerich, G. [Orsay, IAS; AIM, Saclay] (ORCID:0009

Search for ${\text {Z}{}{}} {\text {Z}{}{}} $ and ${\text {Z}{}{}} {\text {H}{}{}} $ production in the ${\text {b}{}{}} {\bar{{\text {b}{}{}}}{}{}} {\text {b}{}{}} {\bar{{\text {b}{}{}}}{}{}} $ final state using proton-proton collisions at $\sqrt{s}=13\,\text {Te}\hspace{-.08em}\text {V} $

A search for ${\text {Z}{}{}} {\text {Z}{}{}} $ and ${\text {Z}{}{}} {\text {H}{}{}} $ production in the ${\text {b}{}{}} {\bar{{\text {b}{}{}}}{}{}} {\text {b}{}{}} {\bar{{\text {b}{}{}}}{}{}} $ final state is presented, where H is the standard model (SM) Higgs boson. The search uses an event sample of proton-proton collisions corresponding to an integrated luminosity of 133$\,\text {fb}^{-1}$ collected at a center-of-mass energy of 13$\,\text {Te}\hspace{-.08em}\text {V}$ with the CMS detector at the CERN LHC. The analysis introduces several novel techniques for deriving and validating a multi-dimensional background model based on control samples in data. A multiclass multivariate classifier customized for the ${\text {b}{}{}} {\bar{{\text {b}{}{}}}{}{}} {\text {b}{}{}} {\bar{{\text {b}{}{}}}{}{}} $ final state is developed to derive the background model and extract the signal. The data are found to be consistent, within uncertainties, with the SM predictions. The observed (expected) upper limits at 95% confidence level are found to be 3.8 (3.8) and 5.0 (2.9) times the SM prediction for the ${\text {Z}{}{}} {\text {Z}{}{}} $ and ${\text {Z}{}{}} {\text {H}{}{}} $ production cross sections, respectively.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS