Engineering PapersSearch

SEARCH · Engineering Papers

Results for “red AI”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

Measuring the Energy Consumption and Efficiency of Deep Neural Networks: An Empirical Analysis and Design Recommendations

Addressing the "Red-AI" trend of rising energy consumption by large-scale neural networks, this study investigates the measured energy consumption of training various fully connected neural network architectures. We introduce the BUTTER-E dataset, an augmentation to the BUTTER Empirical Deep Learning dataset, containing energy consumption and performance data from 41,129 individual experimental runs spanning 30,582 distinct configurations: 13 datasets, 20 sizes (trainable parameters), 8 "shapes", and 14 depths on both CPUs and GPUs using node-level watt-meters. This dataset reveals the complex relationship between dataset size, network structure, and energy use. Our analysis uncovers a surprising, hardware-mediated non-linear relationship between energy efficiency and network design, challenging the assumption that reducing the number of parameters or FLOPs is the best way to achieve greater energy efficiency. We propose a straightforward and effective energy model that accounts for network size, computing, and memory hierarchy. Highlighting the need for cache-considerate algorithm development, we suggest a codesign approach to energy efficient network, algorithm, and hardware design. This work contributes to the fields of sustainable computing and Green AI, offering practical guidance for creating more energy-efficient neural networks and promoting sustainable AI.

97 MATHEMATICS AND COMPUTING

Decision Points and Practical Considerations for AI Projects

In this presentation, I will present business-relevant decisions, risks, and considerations for practical implementations of AI projects. I will use energy efficiency and renewable energy AI projects at NREL as examples and case-studies highlighting the journey from concept to implementation. First, I present challenges, questions, and trade-offs related to system inputs: the data. Next, I will examine issues with system behavior and trust, presenting examples, risks, and mitigation strategies. Finally, I will discuss challenges to effective widespread deployment of AI systems including energy, compute, and time requirements.

AI

Artificial Intelligence for Event Reconstruction and Higgs Physics at CMS and Future Colliders

This dissertation charts a trajectory in which advances in artificial intelligence (AI) play a central role in pushing the high-energy physics frontier, complementing progress driven by higher collision energies and larger colliders. The discovery potential of the LHC and future colliders relies on accurate reconstruction of increasingly complex particle collision events. In the CMS experiment, this task is performed by the particle-flow (PF) algorithm. This dissertation presents the first implementation of a machine-learning-based particle-flow (MLPF) reconstruction in the CMS detector based on transformer architectures. In simulated top quark--antiquark pair (ttbar) events under LHC Run~3 (2023--2024) conditions, MLPF improves jet energy resolution by 10--20\% compared to standard PF for jets with transverse momentum between 30--100\GeV. Runtime performance is evaluated using simulated multijet events, with a median inference time of 20\unit{ms} per event on an NVIDIA L4 GPU, compa red to approximately 110\unit{ms} for standard PF. The MLPF algorithm is also validated on Run~3 collision data, representing the first data-validated ML-based reconstruction pipeline at any LHC experiment. We then extend MLPF toward future electron--positron colliders and introduce the first full-simulation cross-detector transfer learning workflow for PF reconstruction. The model is pre-trained on simulated events from the Compact Linear Collider detector (CLICdet) and fine-tuned on the CLIC-like detector (CLD) proposed for the Future Circular Collider (FCC). This approach achieves up to a 40\% improvement in jet energy resolution over rule-based reconstruction while reducing the required training dataset size by an order of magnitude, demonstrating the potential of AI to accelerate detector development and optimization. This dissertation also demonstrates how modern AI techniques enhance the sensitivity of LHC physics analyses. A CMS search for highly Lorentz-boosted Higgs bosons decaying to \textrm{W} boson pairs is presented, focusing on the single-lepton final state. A dedicated fine-tuning strategy for \ParT yields an approximately 70\% increase in expected sensitivity relative to the baseline model. The analysis uses proton--proton collision data at a center-of-mass energy of \ensuremath{\sqrt{s}=13\TeV} collected by CMS between 2016 and 2018, corresponding to an integrated luminosity of 138\ensuremath{\ \mathrm{fb}^{-1}}. The expected significance of the search is $1.86\sigma$, with an observed signal strength of $-0.19^{+0.48}_{-0.46}$. Finally, explainable AI techniques are applied to the MLPF and \ParticleNet algorithms using layerwise relevance propagation, showing that both models base their predictions on physically meaningful features consistent with our physics intuition. Together, these results demonstrate how advanced AI methods can enhance reconstruction, analysis sensitivity, and interpretability, shaping the next era of experimental parti cle physics.

Mokhtar, Farouk [UC, San Diego]

Digitizing and Enhancing Accessibility of the Fusion Safety Archives

This project focuses on the digitization and public accessibility to the Fusion Safety Archives at the Idaho National Laboratory. The first phase involves a thorough review of each document in the physical archives to determine its online availability. For documents that are available online, PDF copies and unique identifiers are collected for database integration. Documents not available online are delivered to Red Inc. for digitization. Additionally, defunct storage devices such as diskettes are sent to INL’s archival department for data retrieval where possible. The second phase of the project involves the creation of a comprehensive database to house the digital copies of the archives. The database will facilitate easy access and management of the digitized documents. Following the database creation, we plan to train a Retrieval-Augmented Generation (RAG) based AI on publicly available documents. The trained AI will be integrated into a front-facing application, allowing the public to easily access information from the Fusion Safety Archives. This project aims to preserve valuable historical data, improve accessibility, and promote transparency in fusion safety research.

70 - PLASMA PHYSICS AND FUSION TECHNOLOGY

Discovery of tunable and soluble organic emitters for solid-state lasers with a self-driving laboratory

We have recently demonstrated the ability of using self-driving laboratories for AI-driven searches of organic emitters for solid-state lasing devices. Our past workflow featured solubility challenges for such large molecular moieties. In this next-generation study, we return to the drawing board to explore a family of com-pounds that are much solution processable and composed of a set of electronic cores that provide a broader color response. Out of 252 potential candidates,and with guidance from DFT calculations, we selectively perform a compre-hensive study exploring 51 fluorene-based A-B-A type organic laser oligomers, armed with our self-driving lab. The candidates range from simple hydrocarbon molecules to complex hetero atom-mixed molecules. As a result of this study, we highlight diketopyrrolopyrrole and benzodiazole derivatives for their largely red-shifted emissions. Furthermore, we investigate the effect of color change aris-ing from hetero atom permutation, fluorine addition, thiophene coupling, and a combination of fluorine addition and thiophene coupling. Amplified spontaneous emission (ASE) measurements in the solid state further corroborate the lasing potential of selected candidates, reinforcing their suitability for future device applications. The computational study with density functional theory confirms the experimental results.

fluorescence

The BTSbot-nearby Discovery of SN 2024jlf: Rapid, Autonomous Follow-up Probes Interaction in an 18.5 Mpc Type IIP Supernova

We present observations of the Type IIP supernova (SN) SN 2024jlf, including spectroscopy beginning just 0.7 days (∼17 hr) after first light. Rapid follow-up was enabled by the new BTSbot-nearby program, which involves autonomously triggering target-of-opportunity requests for new transients in Zwicky Transient Facility data that are coincident with nearby (D < 60 Mpc) galaxies and identified by the BTSbot machine learning model. Early photometry and nondetections shortly prior to first light show that SN 2024jlf initially brightened by >4 mag day −1 , quicker than ∼90% of Type II SNe. Early spectra reveal weak flash ionization features: narrow, short-lived (1.3 < τ[days] < 1.8) emission lines of Hα, He II , and C IV . Assuming a wind velocity of v w = 50 km s −1 , these properties indicate that the red supergiant progenitor exhibited enhanced mass loss in the last year before explosion. We constrain the mass-loss rate to $1{0}^{-4}\lt \dot{M}\,[{M}_{\odot }\,{\mathrm{yr}}^{-1}]\lt 1{0}^{-3}$ by matching observations to model grids from two independent radiative hydrodynamics codes. BTSbot-nearby automation minimizes spectroscopic follow-up latency, enabling the observation of ephemeral early-time phenomena exhibited by transients.

core-collapse supernovae

Towards AI Based Data Classification for Decision Making During Testing

During the development of high-consequence items, test systems should be capable of differentiating between test failures resulting from narrowly missing requirements versus those indicating potentially catastrophic faults. In many instances, classifying the data corresponds to simply identifying whether measured waveforms have approximately the anticipated shape. Cast in this light, the problem reduces to converting raw data into a form optimal for use with neural network classifiers. This manuscript investigates different means of representing raw data for image classification. Raw data plots and Short Time Fourier Transform (STFT) spectrograms are classified by both custom built, small-scale, Convolution Neural Networks (CNN) and open-source, multi-million parameter, pre-trained deep CNNs. In the case of time varying frequency content, the STFTs provide images with greater detail and can be accurately classified with simpler networks. This requires less memory and runs faster than classifying the raw data using the more sophisticated options—making STFTs optimal for applications with memory constraints. STFTs are not a panacea. In some cases the time-domain signal contains useful information that should not be discarded. Rather than using raw data or STFTs, the images can be constructed from both by using red and green channels of an RGB image to visualize the real and imaginary components of the transform, with the raw data occupying the blue channel.

97 MATHEMATICS AND COMPUTING

High-redshift millennium and astrid galaxies in effective field theory at the field level

Effective field theory (EFT) modeling is expected to be a useful tool in the era of future higher-redshift galaxy surveys such as DESI-II and Spec-S5 due to its robust description of various large-scale structure tracers. However, large values of EFT bias parameters of higher-redshift galaxies could jeopardize the convergence of the perturbative expansion. Here, in this paper we measure the bias parameters and other EFT coefficients from samples of two types of star-forming galaxies in the state-of-the-art MilleniumTNG and astrid hydrodynamical simulations. Our measurements are based on the field-level EFT forward model that allows for precision EFT parameter measurements by virtue of cosmic variance cancellation. Specifically, we consider approximately representative samples of Lyman-break galaxies (LBGs) and Lyman-𝛼 emitters (LAEs) that are consistent with the observed (angular) clustering and number density of these galaxies at 𝑧 = 3. Reproducing the linear biases and number densities observed from existing LAE and LBG data, we find quadratic bias parameters that are roughly consistent with those predicted from the halo model coupled with a simple halo occupation distribution model. We also find nonperturbative velocity contributions (fingers of God) of a similar size for LBGs to the familiar case of luminous red galaxies. However, these contributions are quite small for LAEs despite their large satellite fraction values of up to ∼ 30%. Our results indicate that the effective momentum reach 𝑘 max at 𝑧 = 3 for LAEs (LBGs) will be in the range 0.3−0.6⁢ℎ Mpc −1 (0.2−0.8⁢ℎ Mpc −1 ), suggesting that EFT will perform well for high-redshift galaxy clustering. This work provides the first step toward obtaining realistic simulation-based priors on EFT parameters for LAEs and LBGs.

Sullivan, James M. [Massachusetts Inst. of Technol

A Morphological Model to Separate Resolved–Unresolved Sources in the DESI Legacy Surveys: Application in the LS4 Alert Stream

Separating resolved and unresolved sources in large imaging surveys is a fundamental step to enable downstream science, such as searching for extragalactic transients in wide-field time-domain surveys. Here we present our method to effectively separate point sources from the resolved, extended sources in the Dark Energy Spectroscopic Instrument (DESI) Legacy Surveys (LS). We develop a supervised machine learning model based on the Gradient Boosting algorithm XGBoost. The features input to the model are purely morphological and are derived from the tabulated LS data products. We train the model using ∼2 × 10 5 LS sources in the COSMOS field with HST morphological labels and evaluate the model performance on LS sources with spectroscopic classification from the DESI Data Release 1 (∼2 × 10 7 objects) and the Sloan Digital Sky Survey Data Release 17 (∼3 × 10 6 objects), as well as on ∼2 × 10 8 Gaia stars. A significant fraction of LS sources are not observed in every LS filter, and we therefore build a “Hybrid” model as a linear combination of two XGBoost models, each containing features combining aperture flux measurements from the “blue” (gr) and “red” (iz) filters. The Hybrid model shows a reasonable balance between sensitivity and robustness, and achieves higher accuracy and flexibility compared to the LS morphological typing. With the Hybrid model, we provide classification scores for ∼3 × 10 9 LS sources, making this the largest ever machine learning catalog separating resolved and unresolved sources. The catalog has been incorporated into the real-time pipeline of the La Silla Schmidt Southern Survey (LS4), enabling the identification of extragalactic transients within the LS4 alert stream.

astrostatistics

The ambiguous AT2022rze: changing-look AGN mimicking a supernova in a merging galaxy system

AT2022rze is a luminous, ambiguous transient located south-east of the geometric centre of its host galaxy at redshift $z = 0.08$. The host appears to be formed by a merging galaxy system. The observed characteristics of AT2022rze are reminiscent of active galactic nuclei (AGNs), tidal disruption events, and superluminous supernovae. The transient reached a peak absolute magnitude of $-$20.2 $\pm$ 0.2 mag, showing a sharp rise (t$_{\mathrm{rise,1/e}} = 27.5 \pm 0.6$ d) followed by a slow decline (t$_{\mathrm{dec,1/e}} = 382.9 \pm 0.6$). Its bumpy light curve and narrow Balmer lines indicate the presence of gas (and dust). Its light curve shows rather red colours, indicating that the transient could be affected by significant host extinction. The spectra reveal coronal lines, indicative of high-energy (X-ray/UV) emission. Archival data reveal no prior activity at this location, disfavouring a steady-state AGN, although an optical spectrum obtained prior to the transient is consistent with an AGN classification of the host. Based on this, we conclude that the transient most likely represents a changing-look AGN at the centre of the smallest component of the merging system.

galaxies: active