Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “public release”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Location generalizability of image-based air quality models

This paper is to be submitted at the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) Computer Vision for Earth Observation workshop. The full paper abstract is below: The ability to rapidly quantify atmospheric pollutants is important both for global emissions monitoring and for mitigating the adverse effects that follow a hazardous chemical release. In the aftermath of a chemical release, imagery is often the only available resource to assess local conditions. Recent work has demonstrated initial success in predicting particulate matter pollution from imagery; however, these results are tied to a specific site and do not generalize to new geographic locations. In this work, we seek to understand how easily deep learning models generalize to new locations in the context of image-based air quality assessments, targeting two distinct tasks: (1) broad measures of particulate matter pollution, and (2) the mass of a given chemical released in hazardous plumes. For the latter, we focus on sulfur dioxide, a toxic aerosol and a major component of particulate matter pollution caused by industrial fossil fuel consumption. To develop a model that operates in the widest possible range of environments, we test different training strategies, including the use of new geolocation foundation models. The best performing models achieve >80% accuracy when evaluating unseen imagery at previously seen sites, but we find significant drops in performance when evaluating imagery from unseen sites, at best 65%. Additionally, we present the public release of the National Parks Air Quality Index Dataset, a new medium-sized dataset that pairs imagery with sensor-based air quality measurements at 15 different national parks.

Byler, Eleanor B. [BATTELLE (PACIFIC NW LAB)]↗

HarDWR - Raw Water Rights Records

A dataset within the Harmonized Database of Western U.S. Water Rights (HarDWR). For a detailed description of the database, please see the meta-record v2.0. Changelog v2.0 - Switched source data from collecting records from each state independently to using the WestDAAT dataset v1.0 - Initial public release Description In order to hold a water right in the western United States, an entity, (e.g., an individual, corporation, municipality, sovereign government, or non-profit) must register a physical document with the state's water regulatory agency. State water agencies each maintain their own database containing all registered water right documents within the state, along with relevant metadata such as the point of diversion and place of use of the water. All western U.S. states have digitized their individual water rights databases, as well as geospatial data defining the areas in which water rights are managed. Each state maintains and provides their own water rights data in accordance with individual state regulations and standards. In addition, while all states make their water rights publicly available, each provides their records in unique formats, meaning that file types, field availability, and terms vary from state to state. This leads to additional challenges to managing resources which cross state lines, or conducting consistent multi-state water analyses. For the first version of HarDWR, we collected the water rights databases from 11 Western States of the United States. In order to preform regional analyses with the collected data, the raw records had to be harmonized into one single format. The Water Data Exchange (WaDE) is a program dedicated to the sharing of water-related data for the Western U.S. in a singular consistent format. Created by the Western States Water Council (WSWC) to facilitate the collection and dissemination of water data among WSWC's member states and the public, WaDE provides an important service for those interested in water resource planning and management in their focus region. Of the services which WaDE provides, the one of the most interesting is the WestDAAT dataset, which is a collection of water rights data provided by the 18 WSWC member states that have been standardized into a single format, much like we had done on a more limited scale with HarDWR v1. For this version of HarDWR we decided to use WestDAAT, specifically a snapshot created in Feburary 2024, as our water rights source data. A full explanation of the benefits gained from this switch can be found in the description of the updated Harmonized Water Rights Records v2.0, but in short it has allowed us to focus more of our efforts on answering research questions and gaining a more realistic understanding of how water rights are allocated. For more information on how the data for WestDAAT was collected, please see the WaDE data summary. Terms of Use While WaDE works directly with the state agencies to collect and standardize the water rights records, the ultimate authority for the water rights data remains the individual states. Each state, and their respective water right authorities, have made their water right records available for non-commercial reference uses. In addition, the states make no guarantees as to the completeness, accuracy, or timeliness of their respective databases, let alone the modifications which we, the authors of this paper, have made to the collected records. None of the states should be held liable for using this data outside of its intended use. As several of the states update their water rights databases daily, the information provided here is not the latest possible, and should not be used for legal purposes. WestDAAT itself has irregular updates. Additional questions about the data the source states provided should be directed to the respective state agencies (see methods.csv and organization.csv files described below). In addition, although data was presented here was not collected directly from the states, several states requested specifically worked disclaimers when sharing their data. These disclaimers are included here as an acknowledgement from where the water rights data is primarily sourced. Colorado: "The data made available here has been modified for use from its original source, which is the State of Colorado. THE STATE OF COLORADO MAKES NO REPRESENTATIONS OR WARRANTY AS TO THE COMPLETENESS, ACCURACY, TIMELINESS, OR CONTENT OF ANY DATA MADE AVAILABLE THROUGH THIS SITE. THE STATE OF COLORADO EXPRESSLY DISCLAIMS ALL WARRANTIES, WHETHER EXPRESS OR IMPLIED, INCLUDING ANY IMPLIED WARRANTIES OF MERCHANTABILITY, OR FITNESS FOR A PARTICULAR PURPOSE. The data is subject to change as modifications and updates are complete. It is understood that the information contained in the Web feed is being used at one's own risk." Montana: "The Montana State Library provides this product/service for informational purposes only. The Library did not produce it for, nor is it suitable for legal, engineering, or surveying purposes. Consumers of this information should review or consult the primary data and information sources to ascertain the viability of the information for their purposes. The Library provides these data in good faith but does not represent or warrant its accuracy, adequacy, or completeness. In no event shall the Library be liable for any incorrect results or analysis; any direct, indirect, special, or consequential damages to any party; or any lost profits arising out of or in connection with the use or the inability to use the data or the services provided. The Library makes these data and services available as a convenience to the public, and for no other purpose. The Library reserves the right to change or revise published data and/or services at any time." Oregon: "This product is for informational purposes and may not have been prepared for, or be suitable for legal, engineering, or surveying purposes. Users of this information should review or consult the primary data and information sources to ascertain the usability of the information." File Descriptions The unmodified February, 2024 WestDAAT snapshot is composed of nine files. Below is a brief description of each file, as well as how they were utilized for HarDWR. WaDEDataDictionaryTerms.xlsx: As the file's name implies, this is a data dictionary for all of the below named files. This file describes the column names for each of the following files, with the exception of citation.txt which does not have any columns. The descriptions for each file are divided by tab,with the same name as their associated file, within this document. allocationamount.csv: The "main" file of the group, it contains the water right records for each state. Of particular note, each water right is broken down into one or more water allocations. Allocations may be withdrawn from one or more locations, or even multiple allocations associated with a particular location. This is a more subtle and realistic representation of how water is used than what was available in the first version of HarDWR. For the records from some states, this can mean that multiple allocations listed under a single right will appear as rows within this file. citation.txt: A combination of contact information for WaDE personnel, disclaimer about how the data should be used, and guidelines for citing WestDAAT. methods.csv: A file describing the source and method by which WaDE collected water rights data from each state. organization.csv: A file listing the water rights authoritative agencies for each state. sites.csv: This file provides the geographic, and other descriptors, of the physical location of allocations, called 'sites'. To reiterate, it is possible for one allocation to be associated with multiple sites, as well as one site to be associated with multiple allocations. The two descriptors which we were most interested in where the site's coordinates, as well as whether the site was classified as a Point of Diversion (POD) or a Place of Use (POU). As a general rule, PODs are geographic points, while POUs are areas typically represented as property boundaries or irregularly shaped polygons. sites_pouGeometry.csv: For those allocations with a POU site, this file contains the defining points for the associated polygons. variables.csv: A file describing the units in which an allocation's water amount is reported within WestDAAT. This information is essentially a repeat of the 'AllocationFlow_CFS' and 'AllocationVolume_AF' columns within allocationamount.csv, at least for our purposes. watersources: This file describes the source of water from which each site extracts from. For our purposes, this table was used to determine whether the water came from Surface Water, Groundwater, or Unspecified Water.

Lisk, Matthew↗

A Deep, High-angular-resolution 3D Dust Map of the Southern Galactic Plane

We present a deep, high-angular-resolution 3D dust map of the southern Galactic plane over 239° < l < 6° and ∣ b∣ < 10° built on photometry from the DECaPS2 survey, in combination with photometry from VISTA Variables in the Via Lactea, the Two Micron All Sky Survey, and “Unofficial” Wide-field Infrared Survey Explorer and parallaxes from Gaia Data Release 3 where available. To construct the map, we first infer the distance, extinction, and stellar types of over 700 million stars using the brutus stellar inference framework with a set of theoretical MESA Isochrone and Stellar Tracks (MIST) stellar models. Our resultant 3D dust map has an angular resolution of 1′ , roughly an order of magnitude finer than existing 3D dust maps and comparable to the angular resolution of the Herschel 2D dust emission maps. We detect complexes at the range of distances associated with the Sagittarius-Carina and Scutum-Centaurus arms in the fourth quadrant, as well as more distant structures out to a maximum reliable distance of d ≈ 10 kpc from the Sun. The map is sensitive up to a maximum extinction of roughly A V ≈ 12 mag. We publicly release both the stellar catalog and the 3D dust map, the latter of which can easily be queried via the Python package dustmaps. When combined with the existing Bayestar19 3D dust map of the northern sky, the DECaPS 3D dust map fills in the missing piece of the Galactic plane, enabling extinction corrections over the entire disk ∣b∣ < 10°. Our map serves as a pathfinder for the future of 3D dust mapping in the era of LSST and Roman, targeting regimes accessible with deep optical and near-infrared photometry but often inaccessible with Gaia.

Milky Way galaxy↗

Beyond Fair: Engagement, Data Usability, and Open Community Productivity through the NASA Open Science Data Repository

The FAIR principle (findable, accessible, interoperable, and reusable) governs the storage and sharing of NASA space biology and health data[1]. These guiding principles maximize reuse of data and the reproducibility of scientific findings. The NASA Open Science Data Repository (OSDR; an expansion of NASA GeneLab) was built on the FAIR principles and houses over 500 studies and close to 1000 datasets from decades of space life sciences experiments. OSDR embodies the FAIR principles through data governance that includes mediated, embargoed, and fully open access data. The FAIR data governance principles were recently proposed to be expanded to encompass a FAIREST framework for assessing research data repositories (FAIR + Engagement, Social connections, and Trust)[2]. FAIREST emphasizes the importance of data repositories engaging with the scientific community and gaining the trust of researchers regarding data quality. Trust also refers to the TRUST principles developed for assessment of digital repositories: Transparency, Responsibility, User Focus, Sustainability, Technology[3]. We present the “Open Science for Life in Space” Analysis Working Groups (AWGs) as evidence regarding the power of engagement, social connections, and trust which has enhanced OSDR’s capabilities and productivity. AWG members engage in two main activities. One, members provide feedback on OSDR scientific standards for data ingestion, curation, and reuse (study, subject and assay metadata; processing pipelines; dataset formats and uniformed structures for machine-readability). Two, AWG members collaborate to mine-reuse OSDR data to conduct scientific analysis. With nearly 800 active members, the AWGs have resulted in 32 publications re-using OSDR data and contributed many papers in two major special issues in Cell (2020) and Nature (2024). AWGs also serve as networking groups, facilitate social connections between researchers at all levels of experience, and also have a social online ‘Forum’ used to keep members informed on projects and opportunities. This community-centric, productive, and trustworthy data culture has resulted in a broader effect with international space agencies, academics, and the commercial space sector wanting to submit their data to OSDR. Ten studies of Inspiration 4 data were recently publicly released by OSDR, as were some JAXA human data. Coming up soon in OSDR are data submissions from the European Space Agency, Virgin Galactic PIs, and SpaceX Polaris Dawn. A major benefit of OSDR is the array of standardized and uniformly formatted data (which was developed through AWG member consensus), from which visualization tools, analysis tools, and machine learning models can be built or trained. This talk will cover the Multi-Study Visualization Tool, the Environmental Data Application, RadLab, and a UCSF-NSF funded knowledge graph biomedical health discovery tool ‘SPOKE’ currently being integrated with OSDR. OSDR also provides training programs in bioinformatics and machine learning to improve the scientific community’s awareness of data availability and to boost their ability to perform data analysis. The increasing engagement of the scientific community and the public with technologies powered by artificial intelligence (AI) heightens the need for data analysis to be transparent. The AI for Life in Space initiative leverages the data products provided in OSDR to train AI models, with an emphasis on explainable and trustworthy AI, which would not be possible without FAIR data and metadata. Overall, here we will demonstrate the importance for NASA life sciences data repositories to adhere to the FAIREST framework, by providing examples and success stories from different aspects of OSDR.

data↗

A new step forward in realistic cluster lens mass modelling: analysis of Hubble Frontier Field Cluster Abell S1063 from joint lensing, X-ray, and galaxy kinematics data

We present a new method to simultaneously and self-consistently model the mass distribution of galaxy clusters that combines constraints from strong lensing features, X-ray emission, and galaxy kinematics measurements. We are able to successfully decompose clusters into their collisionless and collisional mass components thanks to the X-ray surface brightness, as well as use the dynamics of cluster members, to obtain more accurate masses exploiting the fundamental plane of elliptical galaxies. Knowledge from all observables is included through a consistent Bayesian approach in the likelihood or in physically motivated priors. We apply this method to the galaxy cluster Abell S1063 and produce a mass model that we publicly release with this paper. The resulting mass distribution presents different ellipticities for the intra-cluster gas and the other large-scale mass components as well as deviation from elliptical symmetry in the main halo. We assess the ability of our method to recover the masses of the different elements of the cluster using a mock cluster based on a simplified version of our Abell S1063 model. Thanks to the wealth of mutliwavelength information provided by the mass model and the detected X-ray emission, we also found evidence for an ongoing merger event with gas sloshing from a smaller infalling structure into the main cluster. In agreement with previous findings, the total mass, gas profile, and gas mass fraction are all consistent with small deviations from the hydrostatic equilibrium. This new mass model for Abell S1063 is publicly available, as the lenstool extension used to construct it.

79 ASTRONOMY AND ASTROPHYSICS↗

The Dark Energy Survey Supernova Program: Light Curves and 5 Yr Data Release

We present griz photometric light curves for the full 5 yr of the Dark Energy Survey Supernova (DES-SN) program, obtained with both forced point-spread function photometry on difference images (DiffImg) performed during survey operations, and scene modelling photometry (SMP) on search images processed after the survey. This release contains 31,636 DiffImg and 19,706 high-quality SMP light curves, the latter of which contain 1635 photometrically classified SNe that pass cosmology quality cuts. This sample spans the largest redshift (z) range ever covered by a single SN survey (0.1 < z < 1.13) and is the largest single sample from a single instrument of SNe ever used for cosmological constraints. We describe in detail the improvements made to obtain the final DES-SN photometry and provide a comparison to what was used in the 3 yr DES-SN spectroscopically confirmed Type Ia SN sample. We also include a comparative analysis of the performance of the SMP photometry with respect to the real-time DiffImg forced photometry and find that SMP photometry is more precise, more accurate, and less sensitive to the host-galaxy surface brightness anomaly. The public release of the light curves and ancillary data can be found at github.com/des-science/DES-SN5YR and doi:10.5281/zenodo.12720777.

79 ASTRONOMY AND ASTROPHYSICS↗

Senteniel-6 Radio Occultation Product Released by NASA GES DISC to Supplement Satellite Remote Sensing Datasets for PBL Study

The NASA Goddard Earth Sciences Data and Information Services Center (GES DISC) curates hyperspectral atmospheric sounder remote sensing and numerical model reanalysis datasets which have been utilized in the Planetary Boundary Layer (PBL) research and applications. The hyperspectral sounder remote-sensing datasets include the Atmospheric Infrared Sounder (AIRS) on the Aqua satellite to the Cross-track Infrared Sounder (CrIS) on Suomi--National Polar- orbiting Partnership (NPP) and National Oceanic and Atmospheric Administration -20 (N NOAA-20)/ Joint Polar-orbiting Satellite System -1 (JPSS-1). The Modern-Era Retrospective analysis for Research and Applications Version 2 (MERRA-2) global reanalysis product provides a data record commencing in 1980. The sounder remote sensing and reanalysis datasets include temperature, water vapor, and trace gas profile down to the PBL, and also have a derived PBL height as well. A nearly 10-year (June 2006 to December 2015) seasonal and annual PBL height climatology dataset from COSMIC Global Navigation Satellite System (GNSS) radio occultation (RO) measurement is also available from the GES DISC. In collaboration with Sentinel-6 Project, the GES DISC is implementing curation activities for GNSS RO products from the Sentinel-6A/Sentinel-6 Michael Freilich satellite launched on November 21, 2020. Sentinel-6A RO products provide refractivity, temperature, and humidity profile with finer vertical resolution, leveraging PBL research and application as a supplement to the hyperspectral sounder remote sensing and reanalysis products. The public release of Sentinel- 6A RO products is scheduled for mid-October of 2021. In this presentation, we will introduce all Senitnel-6A products and services, and demonstrate use cases studying the PBL by combining these products with other GES DISC archived data products.

Feng Ding↗

Neutron computed tomography of B12W and M8N socket sections of the Arecibo telescope

From neutron user principal investigator: We kindly request the public release of three neutron imaging datasets through ONCat. All datasets were collected from two forensic specimens, B12W and M8N, sectioned from zinc-filled steel-wire sockets recovered from the collapsed Arecibo Telescope. The dataset titled “Neutron radiographs of B12W and M8N socket sections of the Arecibo telescope” contains normalized two-dimensional (2D) neutron radiographs of the specimens, showing the geometry and spatial distribution of the steel wires embedded within the zinc matrix, as well as internal features such as voids and cracks. The dataset titled “Neutron computed tomography of B12W and M8N socket sections of the Arecibo telescope” contains normalized 2D neutron projection images acquired over a range of specimen rotation angles for one selected region of each specimen. These projection images were used to reconstruct three-dimensional (3D) tomographic volumes that reveal the embedded-wire geometry and internal defects. The dataset titled “Bragg edge imaging (BEI) of B12W and M8N socket sections of the Arecibo telescope” contains six time-of-flight (TOF) neutron imaging datasets, three from each specimen, acquired at regions of interest selected based on the radiographs. The spatially resolved 2D TOF images show the zinc matrix and embedded steel wires, and the wavelength-dependent neutron transmission data were used to characterize crystallographic texture within the zinc. All components and their condition are in the public domain as they are the property of the National Science Foundation (NSF). The neutron imaging data, part geometries, and detailed forensic information have been widely published in the Arecibo Telescope Collapse Forensic Report by Thornton Tomasetti Engineers and others (NASA report and NASEM report).

Bilheux, Hassina↗

Registration of ‘Independence’ switchgrass

Switchgrass (Panicum virgatum L.), a valuable forage and bioenergy crop, is established more easily than other native perennial warm-season grasses, but its establishment is still slower than that of annual crops. Vigorous switchgrass establishment is crucial for achieving its full potential yield and for effectively competing with weeds for water and nutrient availability. To satisfy this demand, ‘Independence’ (Reg. no. CV-295, PI 704577) switchgrass was developed at the University of Illinois at Urbana-Champaign. Independence was selected for establishment vigor, winter survivorship, and high biomass yield for two cycles from ‘Kanlow’. Here, it is characterized by rapid establishment, robust seedling growth, and the capacity to achieve peak production by the second year. Independence is well adapted to USDA hardiness zones 5b–7b. In field experiments conducted from 2016 to 2017, averaged over seven locations and all years, Independence annually yielded 13 Mg ha –1 of biomass, outperforming ‘Cave-in-Rock’ by 31%, ‘Liberty’ by 15%, ‘Shawnee’ by 42%, ‘Summer’ by 81%, and ‘Sunburst’ by 129%. In wet marginal sites in Illinois from 2020 to 2023, Independence exhibited an average biomass yield of 12 Mg ha –1 , outperforming Shawnee by 31%, Liberty by 27%, and Kanlow by 19%, indicating its potential use on less productive land for annual crops. Independence was publicly released by the University of Illinois at Urbana-Champaign in October 2021.

09 BIOMASS FUELS↗

Genomic variation within the maize stiff-stalk heterotic germplasm pool

The stiff-stalk heterotic group in Maize (Zea mays L.) is an important source of inbreds used in U.S. commercial hybrid production. Founder inbreds B14, B37, B73, and, to a lesser extent, B84, are found in the pedigrees of a majority of commercial seed parent inbred lines. We created high-quality genome assemblies of B84 and four expired Plant Variety Protection (ex-PVP) lines LH145 representing B14, NKH8431 of mixed descent, PHB47 representing B37, and PHJ40, which is a Pioneer Hi-Bred International (PHI) early stiff-stalk type. Sequence was generated using long-read sequencing achieving highly contiguous assemblies of 2.13-2.18 Gbp with N50 scaffold lengths >200 Mbp. Inbred-specific gene annotations were generated using a core five-tissue gene expression atlas, whereas transposable element (TE) annotation was conducted using de novo and homology-directed methodologies. Compared with the reference inbred B73, synteny analyses revealed extensive collinearity across the five stiff-stalk genomes, although unique components of the maize pangenome were detected. Comparison of this set of stiff-stalk inbreds with the original Iowa Stiff Stalk Synthetic breeding population revealed that these inbreds represent only a proportion of variation in the original stiff-stalk pool and there are highly conserved haplotypes in released public and ex-Plant Variety Protection inbreds. Despite the reduction in variation from the original stiff-stalk population, substantial genetic and genomic variation was identified supporting the potential for continued breeding success in this pool. The assemblies described here represent stiff-stalk inbreds that have historical and commercial relevance and provide further insight into the emerging maize pangenome.

59 BASIC BIOLOGICAL SCIENCES↗

Lepton flavor asymmetries: from the early Universe to BBN

Large primordial lepton flavor asymmetries with almost vanishing total baryon-minus-lepton number can evade the usual BBN and CMB constraints if neutrino oscillations lead to perfect flavor equilibration. Solving the momentum averaged quantum kinetic equations (QKEs) describing neutrino oscillations and interactions, we perform the first systematic investigation of this scenario, uncovering a rich flavor structure in stark contradiction to the assumption of simple flavor equilibration. We find (i) a particular direction in flavor space, ∆ne ≃ – 2/3 (– 1)∆n μ for normal (inverted) neutrino mass hierarchy, in which the flavor equilibration is efficient and primordial asymmetries are essentially unconstrained, (ii) a minimal washout factor, ∆$n_{e}^{2}$| BBN ≤ 0.03 (0.016) ∑ α ∆$n_{α}^{2}$| ini yielding a conservative estimate for the allowed primordial asymmetries in a generic flavor direction, and (iii) particularly strong or weak washout if one of the initial flavor asymmetries vanishes due to non-adiabatic muon- or electron-driven MSW transitions. These results open up the possibility of a first-order QCD phase transition facilitated by large lepton asymmetries as well as baryogenesis from large and compensated ∆n e = ∆n μ asymmetries. Our systematic approach of deriving momentum averaged QKEs includes collision terms beyond the damping approximation, energy transfer between the neutrino and electron-photon plasma, and provides a fast and reliable way to investigate the impact of primordial lepton asymmetries at the time of BBN. We publicly release the Mathematica code COFLASY-M on https://github.com/mariofnavarro/COFLASY which solves the QKEs numerically.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

The Surface-Topography Challenge: A Multi-Laboratory Benchmark Study to Advance the Characterization of Topography

Surface performance is critically influenced by topography in virtually all real-world applications. The current standard practice is to describe topography using one of a few industry-standard parameters. The most commonly reported number is Ra, the average absolute deviation of the height from the mean line (at some, not necessarily known or specified, lateral length scale). However, other parameters, particularly those that are scale-dependent, influence surface and interfacial properties; for example the local surface slope is critical for visual appearance, friction, and wear. The present Surface-Topography Challenge was launched to raise awareness for the need of a multi-scale description, but also to assess the reliability of different metrology techniques. In the resulting international collaborative effort, 153 scientists and engineers from 64 research groups and companies across 20 countries characterized statistically equivalent samples from two different surfaces: a “rough” and a “smooth” surface. The results of the 2088 measurements constitute the most comprehensive surface description ever compiled. We find wide disagreement across measurements and techniques when the lateral scale of the measurement is ignored. Consensus is established through scale-dependent parameters while removing data that violates an established resolution criterion and deviates from the majority measurements at each length scale. Our findings suggest best practices for characterizing and specifying topography. The public release of the accumulated data and presented analyses enables global reuse for further scientific investigation and benchmarking.

42 ENGINEERING↗

Automated detection of part quality during two-photon lithography via deep learning

Two-photon lithography (TPL) is an additive manufacturing technique for fabricating three-dimensional objects with nanoscale features. A main challenge of TPL is the routine and labor-intensive task of finding suitable light dosage parameters, i.e. writing speed and laser intensity that induce photo-polymerization within a wide variety of candidate photo-curing polymers. Another challenge is the monitoring required during fabrication. In this work, we apply machine learning (ML) models to accelerate the process of identifying optimal light dosage parameters and automate the detection of part quality. We curate TPL videos of different parts fabricated under a range of light dosage parameters using different resins and train spatial-temporal ML models on this data. Our results show that ML models can detect TPL part quality with a 95.1% accuracy in milliseconds. We also evaluate classification failures and identify two operating modes: parameter optimization and part quality detection. Last but not least, we publicly release this labelled dataset so that it may serve as a useful benchmark to the community. Our approach to process optimization and part quality detection addresses important aspects of TPL industrialization, is applicable beyond TPL and should benefit other additive manufacturing techniques with similar barriers to operating at industrial scale.

36 MATERIALS SCIENCE↗

Gap-filling eddy covariance methane fluxes: Comparison of machine learning model predictions and uncertainties at FLUXNET-CH4 wetlands

Time series of methane fluxes measured by eddy-covariance require gap-filling to estimate annual emissions. Gap-filling methane fluxes is challenging because of high variability and complex responses to multiple drivers. To date, there is no widely established gap-filling standard for methane, with regards both to the best model algorithms and predictors. In this study, we address the need for standardization by synthesizing results of gap-filling methods applied at 17 wetland sites spanning boreal to tropical regions including all major wetlands classes and two rice paddies. We introduce new procedures for: 1) creating realistic artificial gap scenarios, 2) training and evaluating gap-filling models without overstating performance, and 3) predicting half-hourly methane fluxes and annual emissions with robust uncertainty estimates. We tested a conventional method (marginal distribution sampling) and four machine learning algorithms - penalized linear regression, artificial neural networks, random forests, and boosted decision trees - and four predictor sets, including temporal, meteorological, ecosystem carbon and energy flux, and soil predictors. We find that the conventional method can achieve similar median performance to the machine learning models but is worse than the best machine learning models and relatively insensitive to predictor choices. Of the machine learning models, decision tree algorithms performed the best in cross-validation experiments, even with a baseline predictor set, and artificial neural networks showed comparable performance when using all predictors. Soil temperature was frequently the most important predictor whilst water table depth was important at sites with substantial water table fluctuations, highlighting the value of data on soil conditions. Raw gap-filling uncertainties from the machine learning models were underestimated and we propose a method to calibrate uncertainties to observations. Finally, we gap-fill and provide summary evaluation metrics for all 81 sites in the FLUXNET-CH4 community dataset and publicly release the python code for model development, evaluation, and uncertainty estimation.

42 ENGINEERING↗

Semi-Empirical Shadow Molecular Dynamics: A PyTorch Implementation

Here, extended Lagrangian Born–Oppenheimer molecular dynamics (XL-BOMD) in its most recent shadow potential energy version has been implemented in the semiempirical PyTorch-based software PySeQM. The implementation includes finite electronic temperatures, canonical density matrix perturbation theory, and an adaptive Krylov subspace approximation for the integration of the electronic equations of motion within the XL-BOMB approach (KSA-XL-BOMD). The PyTorch implementation leverages the use of GPU and machine learning hardware accelerators for the simulations. The new XL-BOMD formulation allows studying more challenging chemical systems with charge instabilities and low electronic energy gaps. The current public release of PySeQM continues our development of modular architecture for large-scale simulations employing semi-empirical quantum-mechanical treatment. Applied to molecular dynamics, simulation of 840 carbon atoms, one integration time step executes in 4 s on a single Nvidia RTX A6000 GPU.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

A database of thermally activated delayed fluorescent molecules auto-generated from scientific literature with ChemDataExtractor

A database of thermally activated delayed fluorescent (TADF) molecules was automatically generated from the scientific literature. It consists of 25,482 data records with an overall precision of 82%. Among these, 5,349 records have chemical names in the form of SMILES strings which are represented with 91% accuracy; these are grouped in a subsidiary database. Each data record contains one of the following four properties: maximum emission wavelength (λ EM ), photoluminescence quantum yield (PLQY), singlet-triplet energy splitting (ΔE ST ), and delayed lifetime (τ D ). The databases were created through text mining using ChemDataExtractor, a chemistry-aware natural-language-processing toolkit, which has been adapted for TADF research. The text-mined corpus consisted of 2,733 papers from the Royal Society of Chemistry and Elsevier. To the best of our knowledge, these databases are the first databases that have been auto-generated for TADF molecules from existing publications. The databases have been publicly released for experimental and computational applications in the TADF research field.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Planck 2018 results XII. Galactic astrophysics using polarized dust emission

Observations of the submillimetre emission from Galactic dust, in both total intensity I and polarization, have received tremendous interest thanks to the Planck full-sky maps. In this paper we make use of such full-sky maps of dust polarized emission produced from the third public release of Planck data. As the basis for expanding on astrophysical studies of the polarized thermal emission from Galactic dust, we present full-sky maps of the dust polarization fraction p, polarization angle ψ, and dispersion function of polarization angles S . The joint distribution (one-point statistics) of p and NH confirms that the mean and maximum polarization fractions decrease with increasing NH. The uncertainty on the maximum observed polarization fraction, pmax = 22.0 -1.4 +3.5% at 353 GHz and 80' resolution, is dominated by the uncertainty on the Galactic emission zero level in total intensity, in particular towards diffuse lines of sight at high Galactic latitudes. Furthermore, the inverse behaviour between p and S found earlier is seen to be present at high latitudes. This follows the S ∝ p -1 relationship expected from models of the polarized sky (including numerical simulations of magnetohydrodynamical turbulence) that include effects from only the topology of the turbulent magnetic field, but otherwise have uniform alignment and dust properties. Thus, the statistical properties of p, ψ, and S for the most part reflect the structure of the Galactic magnetic field. Nevertheless, we search for potential signatures of varying grain alignment and dust properties. First, we analyse the product map S × p, looking for residual trends. While the polarization fraction p decreases by a factor of 3-4 between NH = 10 20 cm -2 and NH = 2 x 10 22 cm -2 , out of the Galactic plane, this product S × p only decreases by about 25%. Because S is independent of the grain alignment efficiency, this demonstrates that the systematic decrease in p with N H is determined mostly by the magnetic-field structure and not by a drop in grain alignment. This systematic trend is observed both in the diffuse interstellar medium (ISM) and in molecular clouds of the Gould Belt. Second, we look for a dependence of polarization properties on the dust temperature, as we would expect from the radiative alignment torque (RAT) theory. We find no systematic trend of S × p with the dust temperature T d , whether in the diffuse ISM or in the molecular clouds of the Gould Belt. In the diffuse ISM, lines of sight with high polarization fraction p and low polarization angle dispersion S tend, on the contrary, to have colder dust than lines of sight with low p and high S . We also compare the Planck thermal dust polarization with starlight polarization data in the visible at high Galactic latitudes. The agreement in polarization angles is remarkable, and is consistent with what we expect from the noise and the observed dispersion of polarization angles in the visible on the scale of the Planck beam. The two polarization emission-to-extinction ratios, R P/p and R S/V , which primarily characterize dust optical properties, have only a weak dependence on the column density, and converge towards the values previously determined for translucent lines of sight. We also determine an upper limit for the polarization fraction in extinction, p V /E(B - V), of 13% at high Galactic latitude, compatible with the polarization fraction p ≈ 20% observed at 353 GHz. Taken together, these results provide strong constraints for models of Galactic dust in diffuse gas.

79 ASTRONOMY AND ASTROPHYSICS↗