Search NASA⌕ Search

SEARCH · Search NASA

Results for “big data tools”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Emission Lines and the High Energy Continuum

Quasars show many striking relationships between line and continuum radiation whose origins remain a mystery. FeII, [OIII], Hbeta, and HeII emission line properties correlate with high energy continuum properties such as the relative strength of X-ray emission, and X-ray continuum slope. At the same time, the shape of the high energy continuum may vary with luminosity. An important tool for studying global properties of Quasi Stellar Objects (QSOs) is the co-addition of data for samples of QSOS. We use this to show that X-ray bright (XB) QSOs show stronger emission lines in general, but particularly from the narrow line region. The difference in the [OIII]/Hbeta ratio is particularly striking, and even more so when blended FeII emission is properly subtracted. Weaker narrow forbidden lines ([OII] and NeV) are enhanced by factors of 2 to 3 in both UV and optical XB composite spectra. The physical origin of these diverse and interrelated correlations has yet to be determined. Unfortunately, many physically informative trends intrinsic to QSOs may be masked by dispersion in the data due to either low signal-to-noise or variability. An important tool for studying global properties of QSOs is the co-addition of data for samples of QSOS. We use this to show that X-ray bright (XB) QSOs show stronger emission lines in general, but particularly from the narrow line region. The difference in the [OIII]/Hbeta ratio is particularly striking, and even more so when blended Fell emission is properly subtracted. Weaker narrow forbidden lines ([OII] and NeV) are enhanced by factors of 2 to 3 in both UV and optical XB composite spectra. We describe a large-scale effort now underway to probe these effects in large samples, using both data and analysis as homogeneous as possible. Using an HST FOS Atlas of QSO spectra, with primary comparison to ROSAT PSPC spectral constraints, we will model the Big Blue Bump, its relationship to luminosity and QSO type, and we will analyze and contrast line emission and UV/X-ray continuum properties. Absorption of the continuum near the broad emission line region may play a profound role, which we will be able to constrain by direct analysis of observed UV/X-ray spectral absorption.

Green, Paul↗

Digital Technologies at NASA for Science and Engineering

While scientific and engineering advancements used to rely primarily on theoretical studies and physical experiments, today digital technology enabled by petaflops-scale supercomputers is an equal, if not a greater, contributor to such achievements. In addition, computational modeling and simulation serves as a predictive tool that is not otherwise available. As a result, the use of high performance computing is integral to NASA's work in all mission areas such as space exploration, aeronautics, and scientific discovery. But traditional supercomputing alone is not sufficient for all of the space agency's needs. The success of many NASA missions depends on solving complex computing challenges, some of which are NP-hard (decision theory) if using classical solution methods. Quantum computing promises an unprecedented ability to solve such intractable problems by harnessing quantum mechanical effects such as tunneling, superposition, and entanglement. Another disruptive digital technology is neuromorphic computing that uses brain-inspired lessons to generate new architectures that are much more energy efficient, and capable of massive parallel processing and learning in-situ. Finally, with large amounts of observational and computational data sets, the opportunities of big data and data analytics can be leveraged to enable deep learning and knowledge discovery - it's all a massive digital transformation. This talk will be an overview how NASA utilizes digital technologies for its science and engineering efforts.

Biswas, Rupak↗

Design of a high-temperature experiment for evaluating advanced structural materials

This report describes the design of an experiment for evaluating monolithic and composite material specimens in a high-temperature environment and subject to big thermal gradients. The material specimens will be exposed to aerothermal loads that correspond to thermally similar engine operating conditions. Materials evaluated in this study were monolithic nickel alloys and silicon carbide. In addition, composites such as tungsten/copper were evaluated. A facility to provide the test environment has been assembled in the Engine Research Building at the Lewis Research Center. The test section of the facility will permit both regular and Schlieren photography, thermal imaging, and laser Doppler anemometry. The test environment will be products of hydrogen-air combustion at temperatures from about 1200 F to as high as 4000 F. The test chamber pressure will vary up to 60 psia, and the free-stream flow velocity can reach Mach 0.9. The data collected will be used to validate thermal and stress analysis models of the specimen. This process of modeling, testing, and validation is expected to yield enhancements to existing analysis tools and techniques.

Mockler, Theodore T.↗

Performance Analysis of Data Processing in Distributed File Systems with Near Data Processing

In the era of big data, the escalating volume and velocity of data generation pose significant challenges in data processing. Traditional systems like Spark and Hadoop manage the increasing amount and velocity of data by improving data placement and processing speeds. However, they face inherent limitations due to the essential data movement required for processing. In this paper, we explore the Skyhook framework, a novel extension of the Ceph distributed system, which significantly reduces the need for data movement. We present an extensive case study using the Skyhook framework, applying it with the TPC-H and K-means clustering algorithms. More specifically, we leverage the TPC-H benchmark to distinguish between CPU-intensive and I/O-intensive tasks. We explore the integration of K-means clustering into SQL, coupled with a near-data processing system to offload the computational burden of the K-means clustering algorithm to storage nodes. We conduct a comprehensive performance evaluation of distributed data processing applications across three processing approaches: traditional layout (baseline), optimized layout, and near-data processing. Additionally, we introduce the use of the FIO tool to simulate real-world system workloads, enabling the measurement of performance metrics such as average latency and CPU utilization. Our research is a significant advance in understanding how to optimize data processing systems to meet the demands of the modern data landscape.

Hou, Shiyue↗

A Machine-Learning Approach to Assess Aircraft Engine System Performance

Artificial intelligence (AI)/machine learning, and big data are transforming the global business environment. They have become the most disruptive technologies for organizations to improve workplace efficiency and productivity. This work explored the application of machine learning-based predictive analytics that would enable aircraft engine designers to estimate engine system performance quickly during the conceptual design stage. Supervised machine-learning algorithm was employed to study patterns in an existing database of production and research turbofan engines, and built predictive analytics for use in predicting system performance of new turbofan designs. Specifically, the author developed deep-learning analytics to predict turbofan system weight, using turbofan design parameters as the input. The predictive analytics were trained and deployed in Keras, an open-source neural networks API (application program interface) written in Python, with TensorFlow (an open-source artificial AI library developed by Google) serving as the backend engine. The current engine-weight prediction results, together with those for the TSFC (thrust specific fuel consumption) and core-size predictions that were studied previously by the author, show that machine learning-based predictive analytics can be an effective, time-saving tool for aircraft engine design-space exploration during the conceptual design stage. It would enable expeditious identification of the best engine design amongst several candidates.

Michael T Tong↗

Assessing the cumulative effects of nearshore habitat restoration actions for multiple populations of juvenile salmon in Whidbey Basin, Washington: foundation and approach for synthesis and evaluation

Ecosystem restoration is a common tool for re-establishing ecosystem processes, structures, and functions to improve biodiversity and services in coastal and estuarine ecosystems. In the Salish Sea, salmon habitats have been fragmented, reduced in size, and diminished in quality, and the ecosystem processes that form and sustain these habitats have been degraded and disrupted as well. This loss is especially prevalent in estuaries, where up to 90% of former salmon habitat has been lost or compromised. Salmon species are integral to the identities and cultures of people in the Pacific Northwest, yet salmon abundances remain at historic lows, especially in urbanized areas. Recent investments in restoration are creating rearing habitat and repairing lost ecosystem function. However, restoration efforts in this region have largely proceeded at the site scale, with less attention to big-picture thinking regarding how restoration will effectively recover degraded or lost habitats for target species. As a result, no landscape-scale evaluation program exists, and the cumulative benefits of multiple interventions are unknown. We describe innovative methods for science synthesis related to the evaluation of cumulative effects of ecosystem restoration for Pacific salmon, using years of existing, but disparate data. Building from previous work on cumulative effects evaluation and incorporating a hierarchy of hypotheses approach, we propose using causal inference across numerous hypotheses in a framework to assess the cumulative benefits to Pacific salmon from multiple estuarine restoration projects. We present the framework as a method that can be used to address many complex questions and provide examples from the Salish Sea where the approach is being implemented. The framework draws on science synthesis from numerous fields and uses a hierarchy of hypotheses, causal analysis at multiple scales, and a new hierarchy of synthesis for assessing multiple lines of evidence documenting restoration effects on Pacific salmon. We propose causal inference to synthesize dissimilar data streams, in our case, to identify various manifestations of cumulative effects of restoration and benefits to salmon, and to further inform restoration and recovery planning. A unifying framework would allow for the detection of thresholds at which restoration provides measurable improvement and would greatly advance understanding of the effects of restoration on ecosystems.

59 BASIC BIOLOGICAL SCIENCES↗

Future of Fuel Savings

Using automation to free up controllers for more strategic management of air traffic is one approach being studied by NASA as it seeks to boost airspace system capacity and efficiency, thereby saving fuel. Heinz Erzberger, a NASA Ames Research Center senior scientist, says the Advanced Airspace Concept (AAC) has been studied for several years. It could increase efficiency 15% by providing optimal routes that cut airlines direct operating costs. A 25% increase in landings on existing runways could follow an important benefit. AAC is one of the efforts to be reviewed by the Joint Planning and Development Organization, an FAA-led initiative by six federal agencies to redesign the U.S. air transportation system by 2025. The main goal is to triple air traffic capacity within 20 years to avert the sort of gridlock that would make fuel consumption only one of many travel nightmares. The automated system approach would allow aircraft to fly optimal trajectories. A trajectory would be defined in the standard three dimensions and eventually include the fourth, time. The management of air traffic by the data-linked exchange of trajectories would start at high altitude and eventually move down to lower altitudes. The automated concept is an outgrowth of the type of tools developed by NASA for use by FAA controllers in managing traffic flows over the years, including ones that optimize routings for the best fuel burn. But AAC would push automation further to reduce workload so controllers can focus on "solving strategic control problems, managing traffic flow during changing weather and ... other unusal events." One key component, the automated trajectory server (ATS), is a ground systems that would rely on software to manage flight path requests from aircrews and controllers. But, Erzberger acknowledges, "The FAA's current plan for upgrades to air traffic services does not include [allowing] the future ground system to issue separation-critical clearances of trajectory changes autonomously to aircraft via data link without explicit approval of a controller," as the AAC proposes. The AAC enables pilots or controllers to data link requests for a trajectory change to the ATS for approval after they are deconflicted with the paths of other aircraft. To divert around storms, for example, pilots could data link their trajectory preference to the ATS. Since several aircraft might request similar routes, the computer would then have to suggest alternatives. This could be accomplished without pilot-controller radio calls, a big bottleneck now. The ATS would have a built-in conflict monitor to call for a resolution (turn, climb or descend), when loss of separation is likely in 1-20 min. The AAC system would reduce controller errors by 90%, according to NASA Ames estimates. The AAC would have a back-up program to assure separation-Tactical Separation Assurance (TSAFE). It s designed to detect short-term traffic conflicts within 3-4 min. of loss of separation. The last line of defense would still be provided by traffic alert & collision avoidance systems (TCAS).

Hughes, David↗

Digital Lunar Exploration Sites (DLES) Terrain Crafting

Humans will soon be returning to the surface of the Moon with NASA’s Artemis program. The Artemis program is an international collaboration that will consist of a complex series of space systems and missions to explore the lunar surface and pave the way for the future exploration of Mars. NASA and its partners rely heavily on simulation for lighting and navigation studies as well as training astronauts, flight controllers, and mission support staff. The NASA Exploration Systems Simulations (NExSyS) team in the Simulation and Graphics Branch (ER7) in the Engineering Directorate at NASA’s Johnson Space Center has built up many simulation products to support this effort, one of which is the Digital Lunar Exploration Sites (DLES). DLES is a collection of products used to simulate and render the lunar surface in a digital environment. We discussed and presented an overview of the DLES products at the 2022 IEEE Aerospace Conference in Big Sky, MT with a paper titled "Digital Lunar Exploration Sites". This “DLES Terrain Crafting” paper will expand on the information previously provided in “DLES” paper and dive deeper into the details of the terrain crafting process and the toolsets used to support this task. The best digital data currently available of the lunar surface is provided by the Lunar Reconnaissance Orbiter (LRO). Its Lunar Orbiter Laser Altimeter (LOLA) achieves an impressive resolution of 5m per pixel at the Lunar South Pole (LSP) and can generate datasets covering a large continuous region near the LSP. There are a few additional methods, such as Shape from Shading which can infer higher resolution data (up to 1m per pixel) from the LRO Narrow Angle Camera (NAC) images. However, surface-based simulations require higher-resolution data, and this paper will discuss the process of enhancing the terrain to meet that need. The process begins with capturing statistical data of craters in the regions of interest using images provided by the LRO NAC. This data is then used to scatter artificial features which are not captured in the truth data, resulting in an enhanced DEM with a much higher resolution of 20cm per pixel. Many tools were built up to assist in the creation of these artificial Digital Elevation Models (DEM), which this paper will discuss in detail. DEMs themselves are a very powerful representation of a planetary surface, and many operations and tools can utilize the data they contain. This paper includes a description of the rendering of the lunar surface in a graphics engine, generation of contact patches to simulate tire to ground interaction, and ray tracing utilities to model Line of Sight (LOS) interactions with the terrain. This paper will also explore some new tool sets currently under development which aim to utilize Machine Learning (ML) to assist in the identification of craters from LRO NAC imagery. While this is not a novel idea, the NExSyS team is developing a unique approach which may result in more robust identification of crater characteristics.

Artemis↗

Data-Intensive Science meets Inquiry-Driven Pedagogy: Interactive Big Data Exploration, Threshold Concepts, and Liminality

Threshold concepts in any discipline are the core concepts an individual must understand in order to master a discipline. By their very nature, these concepts are troublesome, irreversible, integrative, bounded, discursive, and reconstitutive. Although grasping threshold concepts can be extremely challenging for each learner as s/he moves through stages of cognitive development relative to a given discipline, the learner's grasp of these concepts determines the extent to which s/he is prepared to work competently and creatively within the field itself. The movement of individuals from a state of ignorance of these core concepts to one of mastery occurs not along a linear path but in iterative cycles of knowledge creation and adjustment in liminal spaces - conceptual spaces through which learners move from the vaguest awareness of concepts to mastery, accompanied by understanding of their relevance, connectivity, and usefulness relative to questions and constructs in a given discipline. For example, challenges in the teaching and learning of atmospheric science can be traced to threshold concepts in fluid dynamics. In particular, Dynamic Meteorology is one of the most challenging courses for graduate students and undergraduates majoring in Atmospheric Science. Dynamic Meteorology introduces threshold concepts - those that prove troublesome for the majority of students but that are essential, associated with fundamental relationships between forces and motion in the atmosphere and requiring the application of basic classical statics, dynamics, and thermodynamic principles to the three dimensionally varying atmospheric structure. With the explosive growth of data available in atmospheric science, driven largely by satellite Earth observations and high-resolution numerical simulations, paradigms such as that of dataintensive science have emerged. These paradigm shifts are based on the growing realization that current infrastructure, tools and processes will not allow us to analyze and fully utilize the complex and voluminous data that is being gathered. In this emerging paradigm, the scientific discovery process is driven by knowledge extracted from large volumes of data. In this presentation, we contend that this paradigm naturally lends to inquiry-driven pedagogy where knowledge is discovered through inductive engagement with large volumes of data rather than reached through traditional, deductive, hypothesis-driven analyses. In particular, data-intensive techniques married with an inductive methodology allow for exploration on a scale that is not possible in the traditional classroom with its typical problem sets and static, limited data samples. In addition, we identify existing gaps and possible solutions for addressing the infrastructure and tools as well as a pedagogical framework through which to implement this inductive approach.

Ramachandran, Rahul↗

Reinventing wastewater treatment plants: energy neutral treatment and enhanced fertilizer production through a novel resource recovery center

Wastewater treatment plants (WWTPs) are typically energy intensive, mainly due to the secondary treatment processes such as activated sludge (AS) for treatment of organics as well as nutrients like nitrogen. Nitrogen removal presents a big problem for WWTPs. The main form of nitrogen in wastewater is ammonium, and an AS process uses oxygen to convert ammonium into nitrite and nitrate which is then converted to nitrogen through denitrification process. During anaerobic digestion (AD), organic nitrogen gets degraded, resulting in an effluent stream (centrate) with a high nitrogen content, mostly in the form of ammonium. This contributes 15-30% of total nitrogen to the wastewater influent which further increases energy consumption for aeration. The project aims to transform this conventional municipal WWTPs into energy-neutral, resource-recovering facilities by integrating three core technologies: • Cloth Media Filtration (CMF) to replace conventional primary sedimentation (CPS) and increase the diversion of organics from the energy intensive secondary treatment to AD. This results in reduced energy demand for aeration in the secondary process while simultaneously increasing the biogas production in the anaerobic digesters. • Anerobic Digester to increase biogas and ammonia production. • Membrane Evaporation (ME) to recover ammonia from AD centrate and produce marketable fertilizer. The benefits of proposed WWTP process modifications were evaluated using techno economic analysis (TEA) and life cycle assessment (LCA). For CMF portion of the research a statistical analysis was employed to develop data-driven tools that could be used to enhance and optimize its performance in terms of energy savings and effluent quality. The main objective of this project is to reduce the energy demand for secondary treatment at municipal WWTPs by at least 50%, increase anaerobic digester (AD) biogas and ammonia production by 100% and 120%, respectively, and recover 90% of ammonia from the AD. Integrated CMF, AD, and ME was shown to work synergistically toward achieving these decarbonization targets through energy-positive treatment and fertilizer recovery techniques.

42 ENGINEERING↗

A method for assessing economic, environmental, and reliability tradeoffs of interregional transmission connecting ERCOT (the Texas grid) to the eastern and western grids

Reliable development of the power grid is an evolving concern for humanity due to extreme weather that frequently threatens power sector infrastructure. The state of Texas is a uniquely structured testbed for grid planners to study when looking for solutions to development, innovation, and overcoming such challenges. Because of its size and islanded structure, Texas is small enough to model, but big enough to matter. Texas is a global leader in energy production, energy consumption, and maintains an unusually diverse fuel mix. In addition, the state has experienced winter freezes, heat waves, wind storms, droughts and floods that have threatened power sector infrastructure or caused recent blackouts and calls for demand side conservation. One of the most devastating of these events was the North American winter storm, dubbed “Winter Storm Uri” by the Weather Channel, that froze the region in February 2021 and led to an extended power outage event that put the majority of Texan residents in darkness for days. While preparing to avoid such outage events in the future, various tools have been proposed to improve grid reliability, including energy efficiency, demand response, and distributed energy resources. An additional option would be to develop interregional transmission that connects the Texas grid to other national grids. To assess the merits of this idea, we developed a novel, universally-applicable and internationally-relevant framework to study how the Texas grid would evolve alongside access to various interregional ties. This method allows us to stress the synthetic grid structure and analyze how it would respond to the shock of a simulated winter storm event. Our method leverages open-source modeling tools, such as PowerGenome, pyGRETA, and GenX to synthesize unique zonal grid data, construct a consolidated network of model regions, and simulate different developmental pathways of capacity expansion and operational dispatch. We demonstrate our method with an analysis connecting the Electric Reliability Council of Texas (ERCOT), the grid that serves most of Texas, the Western Electricity Coordinating Council (WECC), the grid that serves the western half of the contiguous U.S., and the Eastern Interconnect, the grid that serves the eastern half of the contiguous U.S. Our results indicate that the cost-optimal capacity of interregional transmission connecting the ERCOT grid to other grids lies between 9–13 GW assuming baseline conditions. Building this amount of connecting capacity in one or multiple directions lowers the costs and emissions of development and operation by up to $16 billion and 257 million metric tonnes (MMT) respectively. Additionally, our results show that the interregional connections between ERCOT and other national grids reduce the amount of total load shed required through mild winter storm events. However, our results also show that there is a threshold of very extreme winter storm conditions, spanning multiple service areas, above which the connections exacerbate resource adequacy problems. Therefore, the results indicate that the connections need to be carefully planned alongside the rest of the grid infrastructure to avoid over-reliance on specific resources or technology options.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Unified Modeling Architecture for Load Management in Extreme Heat: The New York City Case

Integration of renewable resources to meet growing energy demand is becoming a global priority under decarbonization mandates. This study contributes to ongoing efforts on this key subject by assessing the feasibility of using coastal-urban renewable energy resources, namely, offshore wind and rooftop photovoltaic systems, to meet electricity demand of New York City during the intense recent heat wave period of June 2025. A unified modeling framework, based on the urbanized weather research and forecasting model, is used to simulate climate, renewable resources, and energy demand variables. Findings show significant energy load mismatch of approximately 1150 GWh over the month, between the demand and the combined renewable generation outcome. Three storage integration scenarios are analyzed to mitigate the deficits, reducing said deficits by a minimum of approximately 9% over the duration of the month. This study provides a transferable modeling framework tool for evaluating renewable integration in dense urban environments that can be used by grid operators to support grid resilience during extreme heat events.

54 ENVIRONMENTAL SCIENCES↗

Marketing Remote Sensing Data for North Pacific Fisheries Development and Management

Fish poaching, drug trafficking, ocean dumping, and other illegal activities are important problems on the high seas and in national economic zones. The primary thrust of the EOCAP II project, "Marketing Remote Sensing Data for North Pacific Fisheries Development and Management", was to use space-based sensors to improve the effectiveness of marine monitoring, control, and surveillance (MCS). Our initial objectives were to concentrate on the development of MCS tools using Advanced Very High Resolution Radiometry (AVHRR) and Synthetic Aperture Radar (SAR) data. Although we have successfully completed development of an initial version of our SAR-based monitoring tool (OmniVision), project activity has resulted in a much broader application of space-based assets to marine applications. Based in part on work commenced within EOCAP II, a new company, Ocean and Coastal Environmental Sensing, Inc. (OCENS), has been launched and the development of several new software products outside of the MCS arena initiated. One of those products, SeaStation, is near completion with a Fall, 1995 release date. Equity investment in OCENS now totals $70,000-with an additional amount being sought in the first round of financing. One of the pre-eminent objectives of EOCAP II is to make contributions to the US economy and job growth through the expansion of commercial uses of remotely sensed data. OCENS and the software products it is introducing into marine and coastal zone markets responds to this primary object*e. EOCAP II funding leveraged the market and technical know-how of OCENS founders into smart products that benefit marine and coastal zone users. Although technical difficulties and geopolitical shifts damaged the commercial feasibility of initial project objectives, the flexibility of the EOCAP II program now permits long-term business success. This in no small part stems from the fact that the EOCAP program recognizes the realities of small and start-up businesses and does not attempt to force these conditions to fit the apparent needs of big government. Instead, EOCAP works with those who know their market best in order to produce successful products and expanding businesses.

Source record↗

Understanding the Thermal Physics and Metallurgy of Metal Big Area Additive Manufacturing

The research goal of this EPSCoR-DOE partnership is to mitigate defects in parts made using a new type of additive manufacturing (AM) process called metal Big Area Additive Manufacturing (m-BAAM). To realize this goal, the PIs will detect and correct defects in the part as it is being printed by combining fundamental knowledge of the thermal physics and metallurgy of m-BAAM with in-process sensor data. Developed at the DOE-funded Manufacturing Demonstration Facility at Oak Ridge National Laboratory, the m-BAAM process involves one or more robots working together to produce a part by fusing metal wire layer-by-layer using arc welding. The process can print large metal parts such as turbine blades, which is not possible using other AM processes. In addition, m-BAAM production rates are more than ten times faster than other AM processes while requiring one-tenth of the material cost. Despite their potential to become a critical force multiplier in the energy generation industry, m-BAAM parts may fail to print accurately due to retention of heat and uneven cooling. Overheating and anomalous cooling rates in turn can cause inconsistencies in the microstructure, leading to sudden failure when used in safety-critical applications. In other words, flaw formation in m-BAAM parts is governed by the thermal history – intensity and spatial distribution of heat inside the part during printing. The thermal history is a complex function of the part shape and process settings such as welding energy, path taken by the welding torch for deposition (tool path), wire feed rate, among others.

36 MATERIALS SCIENCE↗

Federated Giovanni: A Distributed Web Service for Analysis and Visualization of Remote Sensing Data

The Geospatial Interactive Online Visualization and Analysis Interface (Giovanni) is a popular tool for users of the Goddard Earth Sciences Data and Information Services Center (GES DISC) and has been in use for over a decade. It provides a wide variety of algorithms and visualizations to explore large remote sensing datasets without having to download the data and without having to write readers and visualizers for it. Giovanni is now being extended to enable its capabilities at other data centers within the Earth Observing System Data and Information System (EOSDIS). This Federated Giovanni will allow four other data centers to add and maintain their data within Giovanni on behalf of their user community. Those data centers are the Physical Oceanography Distributed Active Archive Center (PO.DAAC), MODIS Adaptive Processing System (MODAPS), Ocean Biology Processing Group (OBPG), and Land Processes Distributed Active Archive Center (LP DAAC). Three tiers are supported: Tier 1 (GES DISC-hosted) gives the remote data center a data management interface to add and maintain data, which are provided through the Giovanni instance at the GES DISC. Tier 2 packages Giovanni up as a virtual machine for distribution to and deployment by the other data centers. Data variables are shared among data centers by sharing documents from the Solr database that underpins Giovanni's data management capabilities. However, each data center maintains their own instance of Giovanni, exposing the variables of most interest to their user community. Tier 3 is a Shared Source model, in which the data centers cooperate to extend the infrastructure by contributing source code.

Giovanni↗

Leveling-Up for Big-Format Modules

Like Mario grabbing a super mushroom, PV modules just keep getting bigger! As they grow, so do the challenges of handling, installing, and testing them in the field. At NREL, we've embarked on our own New Hope - adapting to this size revolution across our tools, transportation, ergonomics, and field compatibility. Join us as we navigate this galactic expansion and keep PV testing at the cutting edge.

14 SOLAR ENERGY↗

Open Science for Life in Space: Data Sharing and Tools for Knowledge Discovery

The next era in human space exploration is rapidly approaching and will require the use of countermeasures to deep space health hazards. The development of countermeasures (or, there-purposing of existing agents) will be highly dependent on our understanding of basic biological responses to space stressors (e.g. ionizing radiation, altered gravitational fields, altered day-night cycles, confinement, isolation, hostile-closed environments, distance-duration from Earth, exposure to celestial regolith, etc.). The fast-growing array of space biological data, which in the past was simply archived after minimal analysis, holds great potential if it can be reorganized and formatted for Open Science. Organizing the data for such analysis is a challenge because of its diverse nature (molecular, cellular, tissue, imaging, whole organism and behavior). We will discuss here several strategies that NASA's Biological and Physical Science Division has put in place to maximize the return on investment for spaceflight bioscience data. Open Science, as a scientific philosophy, is the concept that the more people who have access to the data, the more knowledge will be gained from it. This guiding principle led NASA to develop GeneLab in 2015. GeneLab houses spaceflight and relevant ground-based multi-omics data, and has grown to ~400 transcriptomatic, proteomic, metabolomic and epigenomic datasets from plant, rodent, small animal, and microbial space experiments. GeneLab provides users with various tools for data analysis and a visualization portal that allows users to interact with gene expression data from space-related 'omics experiments. Open Science is also about building scientific communities, and with this spirit in mind, GeneLab has spawned several Analysis Working Groups (AWGs), comprised of more than 200 volunteer scientists. The AWGs initially provided feedback on the processing pipeline and metadata 'omics standards for GeneLab. Over the last few years, they have become a community-driven science enterprise, engaging in large meta-analysis of GeneLab datasets, resulting in 10 publications (beyond the originally submitted research). Overall, the Open Science nature of GeneLab has resulted in a high degree of data-use, resulting in 40 enabled publications by open data. The enormous success and knowledge gained from GeneLab has led to a collection of sister NASA "Open Science Data Repositories (OSDR)" and research support groups. These include the NASA Ames Life Sciences Data Archive (ALSDA), the NASA Biological Institutional Scientific Collection (NBISC), and the Biospecimen Sharing Program (BSP). All are adopting the GeneLab data architecture system to maximize open-access, find-ability, accessibility, interoperability, and reusability (FAIR). ALSDA collects and curates phenotypic-physiological bioimaging-behavioral data from space and space-relevant non-human experiments, oftentimes coming from the same omics-associated experimental datasets found in GeneLab. Since 2021, a community of ~100 researchers have rallied around ALSDA, to provide feedback in a new ALSDA AWG focused on phenotypic-physiological investigation-sample-assay metadata standards (e.g., Micro-Computed Tomography, Light/Flourescence Microscopy, Western Blot, Flow Cytometry, Novel Object Recognition, Elevated Plus Maze, etc. of ~50 assays collected). These standards are part of a new single point-of-entry data submission portal for all non-human Space Biology and Human Research Program principal investigators, to submit, curate, and share their research data. With open-access space biological data now collected and curated together with rich metadata, and with the potential for linkage to "big data" from the international biological and medical communities (NIH, EBI, etc.), the artificial intelligence and machine learning (AI/ML) era has started for Space Biology.

omics↗

Open Science for Life in Space: Data Sharing and Tools for Knowledge Discovery

The next era in human space exploration is rapidly approaching and will require the use of countermeasures to deep space health hazards. The development of countermeasures (or, the re-purposing of existing agents) will be highly dependent on our understanding of basic biological responses to space stressors (e.g. ionizing radiation, altered gravitational fields, altered day-night cycles, confinement, isolation, hostile-closed environments, distance-duration from Earth, exposure to celestial regolith, etc.). The fast-growing array of space biological data, which in the past was simply archived after minimal analysis, holds great potential if it can be reorganized and formatted for Open Science. Organizing the data for such analysis is a challenge because of its diverse nature (molecular, cellular, tissue, imaging, whole organism and behavior). We will discuss here several strategies that NASA’s Biological and Physical Science Division has put in place to maximize the return on investment for spaceflight bioscience data. Open Science, as a scientific philosophy, is the concept that the more people who have access to the data, the more knowledge will be gained from it. This guiding principle led NASA to develop GeneLab in 2015. GeneLab houses spaceflight and relevant ground-based multi-omics data, and has grown to ~400 transcriptomic, proteomic, metabolomic and epigenomic datasets from plant, rodent, small animal, and microbial space experiments. GeneLab provides users with various tools for data analysis and a visualization portal that allows users to interact with gene expression data from space-related ‘omics experiments. Open Science is also about building scientific communities, and with this spirit in mind, GeneLab has spawned several Analysis Working Groups (AWGs), comprised of more than 200 volunteer scientists. The AWGs initially provided feedback on the processing pipeline and metadata ‘omics standards for GeneLab. Over the last few years, they have become a community-driven science enterprise, engaging in large meta-analysis of GeneLab datasets, resulting in 10 publications (beyond the originally submitted research). Overall, the Open Science nature of GeneLab has resulted in a high degree of data re-use, resulting in 38 additional publications derived from the original 67 publication over the past four years. The enormous success and knowledge gained from GeneLab has led to a collection of sister NASA “Open Science Data Repositories (OSDR)” and research support groups. These include the NASA Ames Life Sciences Data Archive (ALSDA), the NASA Biological Institutional Scientific Collection (NBISC), and the Biospecimen Sharing Program (BSP). All are adopting the GeneLab data architecture system to maximize open-access, find-ability, accessibility, interoperability, and reusability (FAIR). ALSDA collects and curates phenotypic-physiological bioimaging-behavioral data from space and space-relevant non-human experiments, oftentimes coming from the same omics-associated experimental datasets found in GeneLab. Since 2021, a community of ~100 researchers have rallied around ALSDA, to provide feedback in a new ALSDA AWG focused on phenotypic-physiological investigation-sample-assay metadata standards (e.g., Micro-Computed Tomography, Light/Fluorescence Microscopy, Western Blot, Flow Cytometry, Novel Object Recognition, Elevated Plus Maze, etc. of ~50 assays collected). These standards are part of a new single point-of-entry data submission portal for all non-human Space Biology and Human Research Program principal investigators, to submit, curate, and share their research data. With open-access space biological data now collected and curated together with rich metadata, and with the potential for linkage to “big data” from the international biological and medical communities (NIH, EBI, etc.), the artificial intelligence and machine learning (AI/ML) era has started for Space Biology. Several other talks will cover these topics in this conference.

life sciences↗