Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Few Shot Learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

End-to-End Workflow for Machine-Learning-Based Qubit Readout With QICK and hls4ml

In this article, we present an end-to-end workflow for superconducting qubit readout that embeds codesigned neural networks into the quantum instrumentation control kit (QICK). Capitalizing on the custom firmware and software of the QICK platform, which is built on Xilinx radiofrequency system-on-chip field-programmable gate arrays (FPGAs), we aim to leverage machine learning (ML) to address critical challenges in qubit readout accuracy and scalability. The workflow utilizes the hls4ml package and employs quantization-aware training to translate ML models into hardware-efficient FPGA implementations via user-friendly Python application programming interfaces. We experimentally demonstrate the design, optimization, and integration of an ML algorithm for single transmon qubit readout, achieving 96% single-shot fidelity with a latency of 32.25 ns and less than 16% FPGA lookup table resource utilization. Our results offer the community an accessible workflow to advance ML-driven readout and adaptive control in quantum information processing applications.

42 ENGINEERING↗

DNABERT-S: pioneering species differentiation with species-aware DNA embeddings

SUMMARY: We introduce DNABERT-S, a tailored genome model that develops species-aware embeddings to naturally cluster and segregate DNA sequences of different species in the embedding space. Differentiating species from genomic sequences (i.e. DNA and RNA) is vital yet challenging, since many real-world species remain uncharacterized, lacking known genomes for reference. Embedding-based methods are therefore used to differentiate species in an unsupervised manner. DNABERT-S builds upon a pre-trained genome foundation model named DNABERT-2. To encourage effective embeddings to error-prone long-read DNA sequences, we introduce Manifold Instance Mixup (MI-Mix), a contrastive objective that mixes the hidden representations of DNA sequences at randomly selected layers and trains the model to recognize and differentiate these mixed proportions at the output layer. We further enhance it with the proposed Curriculum Contrastive Learning (C2LR) strategy. Empirical results on 28 diverse datasets show DNABERT-S's effectiveness, especially in realistic label-scarce scenarios. For example, it identifies twice more species from a mixture of unlabeled genomic sequences, doubles the Adjusted Rand Index (ARI) in species clustering, and outperforms the top baseline's performance in 10-shot species classification with just a 2-shot training. AVAILABILITY AND IMPLEMENTATION: Model, codes, and data are publically available at https://github.com/MAGICS-LAB/DNABERT_S.

Zhou, Zhihan↗

Lesson Plan Prototype for International Space Station's Interactive Video Education Events

The outreach and education components of the International Space Station Program are creating a number of materials, programs, and activities that educate and inform various groups as to the implementation and purposes of the International Space Station. One of the strategies for disseminating this information to K-12 students involves an electronic class room using state of the art video conferencing technology. K-12 classrooms are able to visit the JSC, via an electronic field trip. Students interact with outreach personnel as they are taken on a tour of ISS mockups. Currently these events can be generally characterized as: Being limited to a one shot events, providing only one opportunity for students to view the ISS mockups; Using a "one to many" mode of communications; Using a transmissive, lecture based method of presenting information; Having student interactions limited to Q&A during the live event; Making limited use of media; and Lacking any formal, performance based, demonstration of learning on the part of students. My project involved developing interactive lessons for K-12 students (specifically 7th grade) that will reflect a 2nd generation design for electronic field trips. The goal of this design will be to create electronic field trips that will: Conform to national education standards; More fully utilize existing information resources; Integrate media into field trip presentations; Make support media accessible to both presenters and students; Challenge students to actively participate in field trip related activities; and Provide students with opportunities to demonstrate learning

Zigon, Thomas↗

Climate model postprocessing

The development of new postprocessing software of the climate modeling group is summarized. Code, test, and perform simulations with global general circulation models are described. The models improve understanding and ability to predict the vagaries of weather and climate. To learn from and utilize the model results, it is necessary to create elaborate postprocessing software to allow analysis of the large volume of data produced. The models produce sigma history tapes. The sigma history records are interpolated to pressure history records, which are written on a pressure history tape. The model results are analyzed on pressure surfaces, with snap shots or time averages.

Abeles, J.↗

A unified large language model–based framework for heterogeneous PV image diagnosis

With advances in imaging technologies, modern photovoltaic (PV) systems generate large volumes of heterogeneous image data, including visible, electroluminescence (EL), and infrared (IR) images. Existing PV image analysis models, particularly deep learning approaches, are typically task-specific and lack cross-modality generalization. To address this limitation, this paper proposes an open-source large language model (LLM)–based unified framework for heterogeneous PV image diagnostics. Through task-aware diagnostic prompting, the framework enables analysis of visible, EL, and IR images within a single pipeline, supporting both zero-shot and few-shot inference and binary and multiclass classification. It is compatible with state-of-the-art multimodal LLMs, including ChatGPT, Gemini, Claude, Qwen, and CLIP. The framework is evaluated on PV module condition classification (clean, soiling, snow, hail, and bird droppings) using visible images, cell crack detection using EL images, and hotspot detection using IR images. GPT-5.1 in few-shot mode achieves the best performance, with classification accuracy exceeding 97.3%. Open-source models such as Qwen and CLIP also deliver competitive results on visible images (around 90% accuracy), though their performance is more limited on EL and IR modalities. On the full ELPV dataset, the framework achieves 83.5% zero-shot accuracy, within 2.8% of the supervised CNN baseline, confirming scalability to larger benchmarks. Practical aspects such as reproducibility, response latency, and confidence estimation are systematically analyzed. The framework operates across PV image modalities without modality- or task-specific training, making it well suited as a rapid pre-screening tool to support downstream detailed diagnostics. A benchmark dataset of diverse labeled PV images is also released.

Li, Baojie↗

Remote Sensing and Fluxes Upscaling for Real-world Impact (Workshop Report)

The "Remote Sensing and Fluxes Upscaling for Real-world Impact" workshop, held on July 9-10, 2024, at Lawrence Berkeley National Lab, was a collaborative effort led by the AmeriFlux Management Project, NEON, and the Carbon Dew Community of Practice. The event brought together over 200 registrants and approximately 100 attendees each day, including leading experts, researchers, and practitioners. The primary focus was on bridging the gap between cutting-edge research and practical applications in environmental monitoring by integrating remote sensing and flux data. Key themes included the importance of site-level measurements for validating remote sensing products, providing nature-based climate solutions, and addressing challenges such as instrument costs and the need for standardized methods. At the regional scale, discussions centered on addressing spatial heterogeneity and using high-resolution remote sensing and machine learning methods to enhance data interpretation. Global scale challenges included data consistency, gap filling, and accurate emission source identification, with opportunities for international collaboration and standardized practices to improve global carbon budget assessments. The workshop emphasized the critical need for integrating data across local, regional, and global scales through explicit scale-matching and developed a workflow for scaling flux data using "straight shot" and "explicit nesting" approaches. The event highlighted the importance of connecting scientific research with real-world applications in carbon, energy, and water management, ensuring that advancements translate into tangible societal benefits. These insights will guide future research, technology transfer, and collaboration, maximizing the potential of environmental fluxes to address real-world challenges.

97 MATHEMATICS AND COMPUTING↗

Online LIBS–ML Framework for Dynamic Characterization of Heterogeneous Waste-Derived Gasification Feedstocks

LIBS−ML framework for real time feedstock characterization during continuous conveyor transport Heterogeneous waste derived feedstocks (e.g., waste coal, biomass and blends) introduce rapid variability in heating value and ash chemistry that affect gasifier operation, yet conventional laboratory characterization techniques are too slow to support proactive control. To address this gap, this study reports on an online, in situ, dynamic characterization framework that couple’s laser-induced breakdown spectroscopy (LIBS) with leakage safe machine learning (ML) regression to deliver real time, decision quality predictions of gasifier relevant properties. A controlled sample matrix spanning two different waste coals, two different biomasses, and engineered blends under two particle size conditions were constructed and benchmarked using standardized laboratory analyses for proximate/ultimate properties and ash composition. LIBS spectra were acquired dynamically as material flowed on a conveyor belt, using high energy 1064 nm laser ablation and shot averaging to improve repeatability and precision. Supervised regression models (multi layer perceptron (MLP) /artificial neural network (ANN), random forest (RF), and support vector regression (SVR)) and an optimized weighted ensemble were trained on emission line feature sets using nested cross validation with Bayesian hyperparameter tuning and validated against an independent hold out set. The proposed LIBS−ML workflow achieves near laboratory predictive fidelity across parametric targets (including higher heating value (HHV), ash content, fixed carbon, sulfur, major ash forming oxides, and initial deformation temperature (IDT)), with the weighted ensemble providing a robust default predictor under dynamic measurement conditions. These results demonstrate a practical pathway for real time feedstock characterization that can enable feedforward adjustments and more resilient gasifier operation for variable quality waste derived fuels.

Biomass↗

Characterization and automated optimization of laser-driven proton beams from converging liquid sheet jet targets

Compact, stable, and versatile laser-driven ion sources hold great promise for applications ranging from medicine to materials science and fundamental physics. While single-shot sources have demonstrated favorable beam properties, including the peak fluxes necessary for several applications, high-repetition-rate operation will be necessary to generate and sustain the high average flux needed for many of the most exciting applications of laser-driven ion sources. Further, to navigate through the high-dimensional space of laser and target parameters toward experimental optima, it is essential to develop ion acceleration platforms compatible with machine learning techniques and capable of autonomous real-time optimization. Here, we present a multi-Hz ion acceleration platform employing a liquid sheet jet target. We characterize the laser-plasma interaction and the laser-driven proton beam across a variety of key parameters governing the interaction using an extensive suite of online diagnostics. We also demonstrate real-time, closed-loop optimization of the ion beam maximum energy by tuning the laser wave front using a Bayesian optimization scheme. This approach increased the maximum proton energy by 11% compared to a manually optimized wave front by enhancing the energy concentration within the laser focal spot, demonstrating the potential for closed-loop optimization schemes to tune future ion accelerators for robust high-repetition-rate operation.

Glenn, G. D. [SLAC National Accelerator Laboratory↗

Benchmarking the performance of uncertainty quantification methods for neural network-based interatomic potentials

Machine-learned interatomic potentials (ML-IAPs) continue to gain popularity as accurate, computationally efficient replacements for traditional, physics-based interatomic potentials and expensive ab initio methods. Uncertainty quantification (UQ) of ML-IAPs is a growing area of research as UQ is critical in many applications of IAPs, such as developing curated datasets, active learning-based data augmentation, self-improving models, and estimating the uncertainty of molecular dynamics simulations. In this paper, we construct and benchmark a series of different neural network potentials (NNPs) with varying network architectures to determine the performance of these models with respect to both the mean and uncertainty calibration error. Each NNP method is specifically designed to predict either epistemic or aleatoric uncertainty with particular focus on the differences in behavior between the epistemic and aleatoric uncertainty estimates. We benchmark these methods using multiple datasets common in the ML-IAP literature. The results show that the aleatoric uncertainty from single-shot model architectures is a competitive alternative to ensemble-based epistemic uncertainty predictions in regions of sufficient data-density. However, in regions where the representative data is sparse, aleatoric uncertainty models tend to overpredict and epistemic methods tend to underpredict the actual model error. We conclude that the type of UQ is crucial when discussing performance of probabilistic model results as different methods have different performance characteristics depending on the regime in which they are evaluated. Therefore, the type of UQ method should be carefully evaluated against both the data characteristics and requirements for the intended application.

97 MATHEMATICS AND COMPUTING↗

Providing Housing, Food and Medical Support for 25,000 Katrina Evacuees with 12 Hours Notice: The Harris County Medical Support of the Superdome Evacuees

Hurricane Katrina was responsible for trapping 25,000 people in the New Orleans Superdome and isolating many others throughout Louisiana and Mississippi. The transport of these evacuees to the Reliant Park (Houston, Texas) used 500 buses each containing about 55 people. Processing the arriving evacuees included addressing their health status and medical needs as follows: an initial triage at disembarkation, a secondary triage in the Reliant Astrodome and Center, and definitive clinical care in the Reliant Arena "Katrina" Clinic. Baylor College of Medicine (BCM) physicians boarded buses and identified the sickest for emergency transport to Harris County Hospital District (HCHD) hospitals. BCM departments represented included pediatrics, family and community medicine, internal medicine, radiology, obstetrics and gynecology, orthopedics, surgery, and psychiatry. Astrodome and Center triage was managed by BCM physicians and staffed by HCHD Nurses and volunteers from Texas and beyond. The Reliant Astrodome, Center and Arena reached peak headcounts of 15,000,4500, and 2500, respectively Most evacuees visiting the triage sites in the Astrodome and Center were treated using "over-the-counter" medications with the remaining being transported to the "Katrina" clinic. The clinic was equipped with a lab, pharmacy, digital X-ray, and ultrasound machines in addition to electronic patient records created using 80 computer terminals. The Katrina clinic saw more than 15,000 patients during 15 days of operations (2,000 on the first full day), administered 10,000 tetanus shots, and filled thousands of prescriptions. At the peak of operations, the clinic saw 150 patients/hour with 25 physicians scheduled for each 12-hour shift. Approximately 900 people were transported to hospital emergency rooms. Within 3 weeks of arriving at the Reliant Park facilities, more than 90% of the families found permanent housing, enrolled children in schools, and found work. Using data obtained from manual and electronic medical records, this presentation will document the major milestones and lessons learned from this extraordinary project to help the Katrina evacuees.

Hamilton, Douglas↗

Pointing stabilization of a 1 Hz high-power laser via machine learning

Abstract High-power lasers are vital for particle acceleration, imaging, fusion and materials processing, requiring precise control and high-energy delivery. Laser plasma accelerators (LPAs) demand laser positional stability at focus to ensure consistent electron beams in applications such as X-ray free-electron lasers and high-energy colliders. Achieving this stability is especially challenging for the low-repetition-rate lasers in current LPAs. We present a machine learning method that predicts and corrects laser pointing instabilities in real-time using a high-frequency pilot beam. By preemptively adjusting a correction mirror, this approach overcomes traditional feedback limits. Demonstrated on the BELLA petawatt laser operating at the terawatt level (30 mJ amplification), our method achieved root mean square pointing stabilization of 0.34 and 0.59 $\unicode{x3bc} \mathrm{rad}$ in the x and y directions, reducing jitter by 65% and 47%, respectively. This is the first successful application of predictive control for shot-to-shot stabilization in low-repetition-rate laser systems, paving the way for full-energy petawatt lasers and transformative advances across science, industry and security.

Amodio, Alessio↗

Stochastic noise can be helpful for variational quantum algorithms

Saddle points constitute a crucial challenge for first-order gradient descent algorithms. In notions of classical machine learning, they are avoided, for example, by means of stochastic gradient descent methods. In this work, we provide evidence that the saddle-points problem can be naturally avoided in variational quantum algorithms by exploiting the presence of stochasticity. We prove convergence guarantees and present practical examples in numerical simulations and on quantum hardware. We argue that the natural stochasticity of variational algorithms can be beneficial for avoiding strict saddle points, i.e., those saddle points with at least one negative Hessian eigenvalue. This insight that some levels of shot noise could help is expected to add a new perspective to notions of near-term variational quantum algorithms. Published by the American Physical Society 2025

Liu, Junyu↗

On the connection between least squares, regularization, and classical shadows

Classical shadows (CS) offer a resource-efficient means to estimate quantum observables, circumventing the need for exhaustive state tomography. Here, we clarify and explore the connection between CS techniques and least squares (LS) and regularized least squares (RLS) methods commonly used in machine learning and data analysis. By formal identification of LS and RLS ``shadows'' completely analogous to those in CS---namely, point estimators calculated from the empirical frequencies of single measurements---we show that both RLS and CS can be viewed as regularizers for the underdetermined regime, replacing the pseudoinverse with invertible alternatives. Through numerical simulations, we evaluate RLS and CS from three distinct angles: the tradeoff in bias and variance, mismatch between the expected and actual measurement distributions, and the interplay between the number of measurements and number of shots per measurement. Compared to CS, RLS attains lower variance at the expense of bias, is robust to distribution mismatch, and is more sensitive to the number of shots for a fixed number of state copies---differences that can be understood from the distinct approaches taken to regularization. Conceptually, our integration of LS, RLS, and CS under a unifying ``shadow'' umbrella aids in advancing the overall picture of CS techniques, while practically our results highlight the tradeoffs intrinsic to these measurement approaches, illuminating the circumstances under which either RLS or CS would be preferred, such as unverified randomness for the former or unbiased estimation for the latter.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Strain Gage Loads Calibration Testing with Airbag Support for the Gulfstream III SubsoniC Research Aircraft Testbed (SCRAT)

This paper describes the design and conduct of the strain-gage load calibration ground test of the SubsoniC Research Aircraft Testbed, Gulfstream III aircraft, and the subsequent data analysis and results. The goal of this effort was to create and validate multi-gage load equations for shear force, bending moment, and torque for two wing measurement stations. For some of the testing the aircraft was supported by three airbags in order to isolate the wing structure from extraneous load inputs through the main landing gear. Thirty-two strain gage bridges were installed on the left wing. Hydraulic loads were applied to the wing lower surface through a total of 16 load zones. Some dead-weight load cases were applied to the upper wing surface using shot bags. Maximum applied loads reached 54,000 lb. Twenty-six load cases were applied with the aircraft resting on its landing gear, and 16 load cases were performed with the aircraft supported by the nose gear and three airbags around the center of gravity. Maximum wing tip deflection reached 17 inches. An assortment of 2, 3, 4, and 5 strain-gage load equations were derived and evaluated against independent check cases. The better load equations had root mean square errors less than 1 percent. Test techniques and lessons learned are discussed.

flight tests↗

Transformer-powered surrogates close the ICF simulation-experiment gap with extremely limited data

Abstract Recent advances in machine learning, specifically transformer architecture, have led to significant advancements in commercial domains. These powerful models have demonstrated superior capability to learn complex relationships and often generalize better to new data and problems. This paper presents a novel transformer-powered approach for enhancing prediction accuracy in multi-modal output scenarios, where sparse experimental data is supplemented with simulation data. The proposed approach integrates transformer-based architecture with a novel graph-based hyper-parameter optimization technique. The resulting system not only effectively reduces simulation bias, but also achieves superior prediction accuracy compared to the prior method. We demonstrate the efficacy of our approach on inertial confinement fusion experiments, where only 10 shots of real-world data are available, as well as synthetic versions of these experiments.

97 MATHEMATICS AND COMPUTING↗

Neural networks for estimation of divertor conditions in DIII-D using C III imaging

Deep learning approaches have been applied to images of C III emission in the lower divertor of DIII-D to develop models for estimating the level of detachment and magnetic configuration (X-point location and strike point radial location). The poloidal distance from the target to the C III emission front is used to represent the level of detachment. The models perform well on a test dataset not used in training, achieving $F_1$ scores as high as 0.99 for detachment state classification and root mean squared error (RMSE) as low as 2cm for front location regression. Predictions for shots with intermittent reattachment are studied, with class activation mapping used to aid in interpretation of the model predictions. Based on the success of these models, a third model was trained to predict the X-point location and strike point radial position from C III images. Though the dataset covers only a small range of possible magnetic configurations, the model shows promising results, achieving RMSE around 1cm for the test data.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Improving the efficiency of learning-based error mitigation

Error mitigation will play an important role in practical applications of near-term noisy quantum computers. Current error mitigation methods typically concentrate on correction quality at the expense of frugality (as measured by the number of additional calls to quantum hardware). To fill the need for highly accurate, yet inexpensive techniques, we introduce an error mitigation scheme that builds on Clifford data regression (CDR). The scheme improves the frugality by carefully choosing the training data and exploiting the symmetries of the problem. We test our approach by correcting long range correlators of the ground state of XY Hamiltonian on IBM Toronto quantum computer. We find that our method is an order of magnitude cheaper while maintaining the same accuracy as the original CDR approach. The efficiency gain enables us to obtain a factor of 10 improvement on the unmitigated results with the total budget as small as 2 ⋅ 10 5 shots. Furthermore, we demonstrate orders of magnitude improvements in frugality for mitigation of energy of the LiH ground state simulated with IBM's Ourense-derived noise model.

97 MATHEMATICS AND COMPUTING↗