Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “materials database”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Dynamic in-context learning with conversational models for data extraction and materials property prediction

The advent of natural language processing and large language models (LLMs) has revolutionized the extraction of data from unstructured scholarly papers. However, ensuring data trustworthiness remains a significant challenge. In this paper, we introduce PropertyExtractor, an open-source tool that leverages advanced conversational LLMs such as Google gemini-pro and OpenAI gpt-4, blends zero-shot with few-shot in-context learning, and employs engineered prompts for the dynamic refinement of structured information hierarchies—enabling autonomous, efficient, scalable, and accurate identification, extraction, and verification of material property data. Our tests on material data demonstrate precision and recall that exceed 95% with an error rate of ∼9%, highlighting the effectiveness and versatility of the toolkit. Finally, databases for 2D material thicknesses, a critical parameter for device integration, and energy bandgap values are developed using PropertyExtractor. In particular, for the thickness database, the rapid evolution of the field has outpaced both experimental measurements and computational methods, creating a significant data gap. Our work addresses this gap and showcases the potential of PropertyExtractor as a reliable and efficient tool for the autonomous generation of various material property databases, advancing the field.

Ekuma, Chinedu E. (ORCID:0000000258527556)↗

Thermochemical Data Fusion Using Graph Representation Learning

Large databases are required for “Big Data” applications in catalysis and materials science. Thermochemical databases can be created by combining data from various sources and by correcting low-fidelity datasets to higher accuracy with minimal computation. To achieve this “data fusion”, thermochemical quantities of interest, calculated at various levels of density functional theory (DFT), need to be mapped to the same, high levels of theory. In this work, a graph theoretical, statistical framework is proposed for such tasks. Subgraph frequencies are shown to provide a natural representation for learning these fusion maps. The maps are linear and are learnt with automated descriptor selection. Using a dataset of as few as ~1% from the QM9 database of 133,885 molecules, these models can predict multiple thermochemical quantities at a higher level of theory with an accuracy of 1 kcal/mol. Here, the method is explainable, generalizable, and provides a diagnostic tool for outlier identification

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

CoRE MOF DB: A curated experimental metal-organic framework database with machine-learned properties for integrated material-process screening

Here, we present an updated version of the Computation-Ready, Experimental (CoRE) Metal-Organic Framework (MOF) database, which includes a curated set of computation-ready MOF crystal structures designed for high-throughput computational materials discovery. Data collection and curation procedures were improved from the previous version to enable more frequent updates in the future. Machine-learning-predicted properties, such as stability metrics and heat capacities, are included in the dataset to streamline screening activities. An updated version of MOFid was developed to provide detailed information on metal nodes, organic linkers, and topologies of an MOF structure. DDEC6 partial atomic charges of MOFs were assigned based on a machine-learning model. Gibbs ensemble Monte Carlo simulations were used to classify the hydrophobicity of MOFs. The finalized dataset was subsequently used to perform integrated material-process screening for various carbon-capture conditions using high-fidelity temperature-swing adsorption (TSA) simulations. Our workflow identified multiple MOF candidates that are predicted to outperform CALF-20 for these applications.

CoRE MOF database↗

Overview of the fusion nuclear science facility, a credible break-in step on the path to fusion energy

The Fusion Nuclear Science Facility (FNSF) is examined here as part of a two step program from ITER to commercial power plants. This first step is considered mandatory to establish the materials and component database in the real fusion in-service environment before proceeding to larger electricity producing facilities. The FNSF can be shown to make tremendous advances beyond ITER, toward a power plant, particularly in plasma duration and fusion nuclear environment. A moderate FNSF is studied in detail, which does not generate net electricity, but does reach the power plant blanket operating temperatures. The full poloidal Dual Coolant Lead Lithium (DCLL) blanket is chosen, with alternates being the Helium Cooled Lead Lithium (HCLL) and Helium Cooled Ceramic Breeder/Pebble Bed (HCCB/PB). Several power plant relevant choices are made in order to follow the philosophy of targeted technologies. Any fusion core component must be qualified by fusion relevant neutron testing and highly integrated non-nuclear testing before it can be installed on the FNSF in order to avoid the high probability of constant failures in a plasma-vacuum system. A range of missions for the FNSF, or any fusion nuclear facility on the path toward fusion power plants, are established and characterized by several metrics. A conservative physics strategy is pursued to accommodate the transition to ultra-long plasma pulses, and parameters are chosen to represent the power plant regime to the extent possible. An operating space is identified, and from this, one point is chosen for further detailed analysis, with R = 4.8 m, a = 1.2 m, IP = 7.9 MA, BT = 7.5 T, βN Gr = 0.9, fBS = 0.52, q95 = 6.0, H98 ∼1.0, and Q = 4.0. The operating space is shown to be robust to parameter variations. A program is established for the FNSF to show how the missions for the facility are met, with a He/H, a DD and 5 DT phases. The facility requires ∼25 years to complete its DT operation, including 7.8 years of neutron production, and the remaining spent on inspections and maintenance. The DD phase is critical to establish the ultra-long plasma pulse lengths. The blanket testing strategy is examined, and shows that many sectors have penetrations for heating and current drive (H/CD), diagnostics, or Test Blanket Modules (TBMs). The hot cell is a critical facility element in order for the FNSF to perform its function of developing the in-service material and component database. The pre-FNSF R&D is laid out in terms of priority topics, with the FNSF phases driving the time-lines for R&D completion. A series of detailed technical assessments of the FNSF operating point are reported in this issue, showing the credibility of such a step, and more detailed emphasis on R&D items to pursue. These include nuclear analysis, thermo-mechanics and thermal-hydraulics, liquid metal thermal hydraulics, transient thermo-mechanics, tritium analysis, maintenance assessment, magnet specification and analysis, materials assessments, core and scrape-off layer (SOL)/divertor plasma examinations.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Initial demonstration of automated fuel performance modeling with 1977 EBR-II metallic fuel pins using BISON code with FIPD and IMIS databases

Using the BISON fuel performance code, simulations were conducted using an automated process to read initial and operating conditions from the Fuels Irradiation and Physics Database (FIPD) and Integral Fast Reactor materials information system (IMIS) database, which contains metallic fuel data from the Experimental Breeder Reactor-II (EBR-II). This work demonstrates use of an integrated framework to access the vast majority of EBR-II experimental fuel pin data to support rapid development of fuel performance models for next-generation metallic fuel systems. With this capability, validation for fuel qualification can be performed rapidly. Between IMIS and FIPD, there is enough information to conduct 1977 unique EBR-II metallic fuel pin histories from 24 different experiments, at varying levels of detail between the two databases. Each of these histories includes a high-resolution power history, flux history, coolant channel flow rates, and coolant channel temperatures. Fission gas release (FGR), cumulative damage fraction (CDF), fuel axial swelling, cladding profilometry, and burnup were all simulated in BISON. The results were compared to post-irradiation examination (PIE) results for the initial demonstration of automated BISON modeling. BISON simulations conducted with IMIS and FIPD were in rough agreement with PIE measurements and calculations. Cladding profilometry, FGR, and fuel axial swelling were found to be in rough agreement with PIE measurements, depending on the physics used within the BISON input files. Here, the mechanical contact solver chosen was found to significantly impact axial fuel swelling and cladding strain predictions. CDF values were assessed to see whether pin failure may have been predicted (CDF ≥ 1). This work suggests that continued development of an automated tool for BISON should focus on inclusion of the Fast Flux Test Facility (FFTF) experimental data for a larger database for metallic fuel, improved physical models to better capture fuel performance, such as fuel-cladding interactions, and a more detailed comparison with available PIE data to further the BISON model development.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

A Curated Experimental Compilation Analyzed by Theory Is More than a Review

Macromolecules is an exceptional resource in the field of polymer science and now publishes more than 1000 original articles a year that set the standard for scientific rigor and creative insights. Over the years, these individual contributions have combined to build the foundation of polymer science, broadly and inclusively defined. In addition to the individual articles, many of which are being celebrated in this series of editorials, Macromolecules has published invaluable reviews and perspectives. These scholarly contributions integrate the insights and results from numerous sources into a unified whole and often recommend future directions for the field. Novices and experts alike benefit from these works that capture topics from emerging discoveries to long-pondered topics and everything in between. To explore the importance of Macromolecules’ reviews and perspectives, we considered their influence on the field and found the 1994 review by Fetters et al. entitled “Connection between Polymer Molecular Weight, Density, Chain Dimensions, and Melt Viscoelastic Properties”1 to be a singularity. This review expertly curates and compiles a trove of data to build robust correlations between molecular characteristics and macroscopic viscoelastic properties of polymer melts, in the context of the tube model of entanglements.

36 MATERIALS SCIENCE↗

High-throughput search for magnetic topological materials using spin-orbit spillage, machine learning, and experiments

Magnetic topological insulators and semi-metals have a variety of properties that make them attractive for applications including spintronics and quantum computation. Here, we use systematic high-throughput density functional theory calculations to identify magnetic topological materials from the ≈ 40000 three-dimensional materials in the JARVIS-DFT database. First, we screen materials with net magnetic moment > 0.5 μB and spin-orbit spillage > 0.25, resulting in 25 insulating and 564 metallic candidates. The spillage acts as a signature of spin-orbit induced band-inversion. Then, we carry out calculations of Wannier charge centers, Chern numbers, anomalous Hall conductivities, surface bandstructures, and Fermi-surfaces to determine interesting topological characteristics of the screened compounds. We also train machine learning models for predicting the spillage, bandgaps, and magnetic moments of new compounds, to further accelerate the screening process. We experimentally synthesize and characterize a few candidate materials to support our theoretical predictions.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Evaluation of BISON metallic fuel performance modeling against experimental measurements within FIPD and IMIS databases

Simulations were conducted using the BISON fuel performance code on an automated process to read initial and operating conditions from two databases—the Fuels Irradiation and Physics Database (FIPD) and Integral Fast Reactor Materials Information System (IMIS) database. These databases contain metallic fuel data from the Experimental Breeder Reactor-II (EBR-II) and the Fast Flux Test Facility (FFTF). The work demonstrates use of an integrated framework to access EBR-II fuel pin data for evaluating fuel performance models contained within BISON to predict fuel performance of next-generation metallic fuel systems. Between IMIS and FIPD, there is enough information to conduct 1,977 unique EBR-II metallic fuel pin histories from 29 different experiments, and 338 pins from FFTF MFF-3 and MFF-5 with varying levels of details between the two databases. Each of these fuel performance histories includes a high-resolution power history, flux history, coolant channel flow rates, and coolant channel temperatures, and new model developments in BISON since the initial demonstration of this integrated framework. Fission gas release (FGR), cumulative damage fraction, fuel axial swelling, FCCI wastage thickness, cladding profilometry, and burnup were all simulated in BISON and compared to post-irradiation examination (PIE) results to evaluate BISON fuel performance modeling. Implementation of new fuel performance models into a generic BISON input file coupled with IMIS and FIPD yielded results with a better representation of physics than the initial evaluation of the integrated framework. Cladding profilometry, FGR, and fuel axial swelling were found to be in good agreement with PIE measurements for most of the pins simulated. The chosen mechanical contact solver was found to significantly impact the axial fuel swelling and cladding strain predictions when used in conjunction with the U-Pu-Zr hot-pressing model since it bound the fuel to prevent further swelling and increased hydrostatic stresses. This work suggests that fuel performance modeling in BISON under steady-state conditions represents the PIE data well and should be reassessed when new PIE data become available in IMIS and FIPD databases and when improved physical models to better capture fuel performance are added to BISON.

Paaren, Kyle M.↗

MOFX-DB: An Online Database of Computational Adsorption Data for Nanoporous Materials

Machine learning and data mining coupled with molecular modeling have become powerful tools for materials discovery. Metal-organic frameworks (MOFs) are a rich area for this due to their modular construction and numerous applications. Here, we make data from several previous large-scale studies in MOFs and zeolites from our groups (and new data for N 2 and Ar adsorption in MOFs) easily accessible in one place. The database includes over 3 million simulated adsorption data points for H 2 , CH 4 , CO 2 , Xe, Kr, Ar, and N 2 in over 160 000 MOFs and zeolites, textural properties like pore sizes and surface areas, and the structure file for each material. We include metadata about the Monte Carlo simulations to enable reproducibility. The database is searchable by MOF properties, and the data are stored in a standardized JSON format that that is interoperable with the NIST adsorption database. We also identify several MOFs that meet high performance targets for multiple applications, such as high storage capacity for both hydrogen and methane or high CO 2 capacity plus good Xe/Kr selectivity. Here, by providing this data publicly, we hope to facilitate machine learning studies on these materials, leading to new insights on adsorption in MOFs and zeolites.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Development of Porous Crystalline Materials for Selective Binding of O 2 from Air

Porous materials that take advantage of different chemisorptive behaviors of O 2 and N 2 could be an effective alternative to approaches based on physisorption for this important separation. Recently, the tetra-nuclear cobalt complex [(Co(III) 2 (bpbp)O 2 ) 2 bdc](PF 6 ) 4 (CSD code: GAMVIB; bpbp – = 2,6-bis(N,N-bis(2-pyridylmethyl)aminomethyl)-4-tert-butylphenolato; bdc 2– = 1,4-benzenedicarboxylato) was shown to have potential for O 2 /N 2 separations based on this concept. This observation raises the question of what other known materials have similar properties. Here, we combine structure screening with a high-throughput periodic density functional theory (DFT) workflow to investigate O 2 and N 2 adsorption in materials from validated crystal structure databases (e.g., CoRE-MOF database and CSD database). These calculations identify multiple materials that have similar di-Co clusters to GAMVIB that are predicted to selectively bind O 2 over N 2 , and suggest design rules that can be used to tune the O 2 and N 2 affinities in materials of this kind.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Discovering Dinuclear Dioxygen-Bridged Cobalt(III) Complexes for Selective Binding of O 2 from Air

The design and development of dioxygen activation in porous crystalline materials is a useful avenue for exploring selective adsorption of O 2 that shows significant potential to enable separations of O 2 and N 2 from air. Porous materials featuring redox-active metal centers have received attention regarding selective O 2 adsorption via chemisorption. Drawing inspiration from a dinuclear cobalt material ([(Co(III) 2 (bpbp)O 2 ) 2 bdc](PF 6 ) 4 (CSD code: GAMVIB; bpbp – = 2,6-bis(N,N-bis(2-pyridylmethyl)aminomethyl)-4-tert-butylphenolato; bdc 2– = 1,4-benzenedicarboxylato)) that displays reversible and selective O 2 adsorption, we focus on searching for potential O 2 -selective materials with dinuclear cobalt clusters that have dioxygen-bridged Co(III) complexes. We combine structure screening with a high-level hybrid periodic density functional theory (DFT) workflow to investigate O 2 and N 2 adsorption in materials from validated crystal structure databases (e.g., the CSD database). These calculations identify multiple materials that are predicted to have superior O 2 binding capability relative to GAMVIB. Grand Canonical Monte Carlo (GCMC) simulations based on DFT-developed force fields were performed for selected candidates to estimate the adsorption performance for O 2 /N 2 mixtures.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Computational toolkit for predicting thickness of 2D materials using machine learning and autogenerated dataset by large language model

The thickness of 2D materials not only plays a crucial role in determining the performance of nanoelectronic and optoelectronic devices but also introduces complexities in predicting volume-dependent properties, such as energy storage capacity, due to the intrinsic vacuum within these materials. Although a plethora of experimental techniques, including but not limited to optical contrast, Raman spectroscopy, nonlinear optical spectroscopy, near-field optical imaging, and hyperspectral imaging, facilitate the measurement of 2D material thickness, comprehensive data for many materials remain elusive. Over the past decade, the exponential proliferation of 2D materials and their heterostructures has outstripped the capabilities of conventional experimental and computational approaches. In this evolving landscape, machine learning (ML) has emerged as an indispensable tool, offering a scalable approach to augment these traditional methodologies. Addressing the critical gap, we introduce THICK2D—Thickness Hierarchy Inference and Calculation Kit for 2D Materials. This Python-based computational framework harnesses an autogenerated thickness database, developed using large language models, and advanced ML algorithms to facilitate the rapid and scalable estimation of material thickness, relying solely on crystallographic data. To demonstrate the utility and robustness of THICK2D, we successfully used the toolkit to predict the thickness of more than 8000 2D-based materials, sourced from two extensive 2D materials databases. THICK2D is disseminated as an open-source utility, accessible on GitHub at https://github.com/gmp007/THICK2D, and archived on Zenodo at https://10.5281/zenodo.11216648.

Ekuma, Chinedu E. (ORCID:0000000258527556)↗

Trust Not Verify? The Critical Need for Data Curation Standards in Materials Informatics

The importance of data curation has been recognized in multiple areas of research; however, the discussion of this important issue is only beginning to emerge in materials science. In this Perspective, we highlight the benefits of using the standardized data curation protocols in materials science and discuss current gaps in accurate and reproducible data reporting using case studies drawn from high-impact materials science papers and well-known databases such as the Crystallography Open Database (COD) and the Cambridge Structural Database (CSD). We argue that both experimental and computational materials scientists need to embrace a culture of rigorous data curation as part of modern research data management. We propose a sample data curation pipeline for materials chemistry and illustrate its use by creating two new materials chemistry databases. Here, we hope that this perspective will serve to catalyze further discussion and promote the continuous development of rigorous data curation practices within the materials science research community. We posit that adherence to best practices of data curation will promote and enhance the reliability, reproducibility, and integrity of materials research and enable the development of reliable AI and machine learning models that critically depend on the use of quality data.

Chemical structure↗

The Intermetallic Reactivity Database: Compiling Chemical Pressure and Electronic Metrics toward Materials Design and Discovery

Here, the advent of high-throughput Density Functional Theory (DFT) calculations has supported the creation of large databases containing the quantitative output necessary for constructing theoretical phase diagrams and predicting physical properties. In this Article, we present a complementary resource, the Intermetallic Reactivity Database (IRD), focused on the chemical bonding features of solid-state structures and indicators of potential structural transformations. Each IRD entry augments common features, such as band structures and density of states (DOS) distributions, with chemically motivated information including DFT-Chemical Pressure (CP) schemes and visualizable representations of the atomic charges. Together, these data types enable the rationalization and prediction of potential structural phenomena encountered in intermetallic chemistry, as we illustrate with four examples: the origins of the Y 2 Ni 2 Mg structure in terms of CP features of its parent structures, the anticipation of intergrowth phases from the net atomic CPs collected in Al-containing binary phases, the correlation between trends in the CP schemes of CaCu 5 -type phases and experimentally observed structural variations, and finally, the development of theoretical methodology with the testing of a streamlined method generating DFT-CP schemes. Altogether, these examples highlight how the IRD supports the creation of models of structural chemistry that extend beyond the bounds of its entries.

36 MATERIALS SCIENCE↗

Detection of topological materials with machine learning

Databases compiled using ab initio and symmetry-based calculations now contain tens of thousands of topological insulators and topological semimetals. This makes the application of modern machine learning methods to topological materials possible. Using gradient boosted trees, we show how to construct a machine learning model which can predict the topology of a given existent material with an accuracy of 90%. Such predictions are orders of magnitude faster than actual ab initio calculations. In this work, we use machine learning models to probe how different material properties affect topological features. Notably, we observe that topology is mostly determined by the “coarse-grained” chemical composition and crystal symmetry and depends little on the particular positions of atoms in the crystal lattice. We identify the sources of our model's errors and we discuss approaches to overcome them.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Prototype crystal structures for exa-AMD framework

This dataset provides a curated collection of initial crystal structure prototypes for the exa-AMD framework, serving as templates for generating hypothetical candidates in multinary materials discovery through elemental substitution and scaling. It includes 36553 ternary and 5254 quaternary unique structures sourced from public databases such as Materials Project, GNoME, OQMD, AFLOW, and NovoMag, formatted in CIF.

exa-AMD↗

CatTestHub: A benchmarking database of experimental heterogeneous catalysis for evaluating advanced materials

The ability to quantitatively compare newly evolving catalytic materials and technologies is hindered by the widespread availability of catalytic data collected in a consistent manner. While certain catalytic chemistries have been widely studied across decades of scientific research, quantitative comparisons based on literature information is hindered by variability in reaction conditions, types of reported data, and reporting procedures. Here, we present CatTestHub, an open-access database dedicated to benchmarking experimental heterogeneous catalysis data. Combining systematically reported catalytic activity data for selected probe chemistries, with relevant material characterization and reactor configuration information, the database provides a collection of catalytic benchmarks for distinct classes of active site functionality. Through key choices in data access, availability, and traceability, CatTestHub seeks to balance the fundamental information needs of chemical catalysis and the FAIR data design principles. Details of the database architecture and the means through which to navigate it are presented, highlighting examples of catalytic insights readily drawn from the available benchmarking data. In its current iteration, CatTestHub spans over 250 unique experimental data points, collected over 24 solid catalysts, that facilitated the turnover of 3 distinct catalytic chemistries. Here, a roadmap is presented through which to expand the open-access platform that serves as a community wide benchmark, primarily through continuous addition of kinetic information on select catalytic systems by members of the heterogeneous catalysis community at large.

Benchmark↗

Machine Learned Force Field Modeling of Metal Organic Frameworks for CO2 Direct Air Capture

Metal organic frameworks (MOFs) are a large class of porous materials and have garnered significant interest due to their large surface areas and their tunable physical and chemical properties. Numerous prior studies have been performed to screen large databases of this material class for promising DAC sorbent materials. These studies have often relied on classical model potentials. While density functional theory (DFT) calculations have been shown to be very accurate for modeling the interaction of CO2 with MOFs, such calculations are too computationally demanding for statistically significant adsorption predictions. To overcome this barrier, we developed methods for training models to achieve DFT-level accuracy for the forces and energies associated with MOF flexibility and CO2 adsorption using machine learned force fields (MLFFs). These methods were parametrized based on DFT calculations of CO2 in a flexible MOF and used to predict MOF structural properties as well as CO2 adsorption in several MOFs.

Findley, John↗