Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “fast optimization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Revealing the Brønsted-Evans-Polanyi relation in halide-activated fast MoS 2 growth toward millimeter-sized 2D crystals

Achieving large-size two-dimensional (2D) crystals is key to fully exploiting their remarkable functionalities and application potentials. Chemical vapor deposition growth of 2D semiconductors such as monolayer MoS 2 has been reported to be activated by halide salts, for which various investigations have been conducted to understand the underlying mechanism from different aspects. Here, we provide experimental evidence showing that the MoS 2 growth dynamics are halogen dependent through the Brønsted-Evans-Polanyi relation, based on which we build a growth model by considering MoS 2 edge passivation by halogens, and theoretically reproduce the trend of our experimental observations. These mechanistic understandings enable us to further optimize the fast growth of MoS2 and reach record-large domain sizes that should facilitate practical applications.

42 ENGINEERING↗

Develop a Fast Analysis Solver for Welding Sequence Optimization

During the shipbuilding manufacturing process, materials are exposed to significant stresses, as induced both thermally and mechanically, that alter the intended design and significantly affect the production schedule, labor hours (fitting, welding, rework, etc.), and material structural performance. The type and magnitude of deformation of a given structure depends on many factors such as the material, thickness and quality of components, the process heat input, preheat and inter-pass temperatures, type and size of welds, welding sequence and direction, location, sequence, and degree of fixturing. Numerical simulations using finite element analysis (FEA) have long been used to analyze welding-induced structural distortion. For large assemblies, transient thermal elastic-plastic analysis (TEPA) can take days or weeks to run, and optimization of welding sequence is not feasible. Simplified analysis methods were developed to reduce computational time. However, it is challenging to use these techniques to fully optimize welding sequencing because of their applied simplifications in modeling weld details. A fast analysis solver that could be used by the shipbuilding industry is being developed for optimizing welding sequences by taking full advantage of modern GPU-based HPC hardware and incorporating patented acceleration schemes. The accelerated processing factors are up to 2200 times greater for large, multi-pass welded structures.

Yang, Yu-Ping↗

Electrolyte Design for Fast‐Charging Lithium‐Based Batteries

Fast charging is essential for the widespread adoption of lithium (Li)-ion batteries, but it is fundamentally limited by sluggish interfacial kinetics, Li plating, and electrolyte instability at high current densities. Over the past decade, electrolyte engineering has emerged as a key strategy to address these challenges. This review summarizes the development of fast-charging electrolytes over the past ten years and outlines a design framework. Electrolyte formulations are first deconstructed into their main components—solvents, salts, and functional additives—and representative strategies for tuning solvation structure and interphase chemistry are discussed to suppress Li plating and improve interfacial kinetics. The discussion then extends to advanced electrolyte systems, particularly localized high-concentration electrolytes (LHCEs), and their compatibility with different anode chemistries. Advanced characterization techniques are also summarized and categorized based on destructiveness, spatial and temporal resolution, quantitative analysis, and the chemical species or processes probed across multiple length scales. Recent progress in AI-enabled electrolyte discovery and battery management system (BMS) strategies for optimized fast-charging protocols is further highlighted. Finally, perspectives are presented on translating electrolyte innovations from academic research to practical applications, with emphasis on cell format, realistic operating conditions, and manufacturability.

25 ENERGY STORAGE↗

Fast Local Spatial Verification for Feature-Agnostic Large-Scale Image Retrieval

Images from social media can reflect diverse viewpoints, heated arguments, and expressions of creativity, adding new complexity to retrieval tasks. Researchers working on Content-Based Image Retrieval (CBIR) have traditionally tuned their algorithms to match filtered results with user search intent. However, we are now bombarded with composite images of unknown origin, authenticity, and even meaning. With such uncertainty, users may not have an initial idea of what the search query results should look like. For instance, hidden people, spliced objects, and subtly altered scenes can be difficult for a user to detect initially in a meme image, but may contribute significantly to its composition. It is pertinent to design systems that retrieve images with these nuanced relationships in addition to providing more traditional results, such as duplicates and near-duplicates — and to do so with enough efficiency at large scale. In this work, we propose a new approach for spatial verification that aims at modeling object-level regions using image keypoints retrieved from an image index, which is then used to accurately weight small contributing objects within the results, without the need for costly object detection steps. We call this method the Objects in Scene to Objects in Scene (OS2OS) score, and it is optimized for fast matrix operations, which can run quickly on either CPUs or GPUs. It performs comparably to state-of-the-art methods on classic CBIR problems (Oxford 5K, Paris 6K, and Google-Landmarks), and outperforms them in emerging retrieval tasks such as image composite matching in the NIST MFC2018 dataset and meme-style imagery from Reddit.

42 ENGINEERING↗

Neural net modeling of equilibria in NSTX-U

Neural networks (NNs) offer a path towards synthesizing and interpreting data on faster timescales than traditional physics-informed computational models. In this work we develop two NNs relevant to equilibrium and shape control modeling, which are part of a suite of tools being developed for the National Spherical Torus Experiment-Upgrade for fast prediction, optimization, and visualization of plasma scenarios. The networks include Eqnet, a free-boundary equilibrium solver trained on the EFIT01 (Equilibrium FITtting 01) reconstruction algorithm, and Pertnet, which is trained on the Gspert code and predicts the non-rigid plasma response, a nonlinear term that arises in shape control modeling. The NNs are trained with different combinations of inputs and outputs in order to offer flexibility in use cases. In particular, Eqnet can use magnetic diagnostics as inputs and act as an EFIT-like reconstruction algorithm, or, by using pressure and current profile information the NN can act as a forward Grad–Shafranov equilibrium solver. This forward-mode version is envisioned to be implemented in the suite of tools for simulation of plasma scenarios. The reconstruction-mode version gives some performance improvements compared to the online reconstruction code real-time EFIT, especially when vessel eddy currents are significant. Here, we report strong performance for all NNs indicating that the models could reliably be used within closed-loop simulations or other applications. Some limitations are discussed.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Serial-Refine Method for Fast Wake-Steering Yaw Optimization

In this paper we present the Serial-Refine method for quickly finding the optimal yaw angles in wake steering. The method optimizes turbine angles serially from upstream to downstream using a small number of candidate angles. The presented results show that Serial-Refine finds solutions that are at least as good as former conventional optimization approaches but that require much less computation time.

17 WIND ENERGY↗

MuyGPyS

MuyGPs is a GP estimation method that affords fast hyperparameter optimization by way of performing leave-one-out cross-validation. MuyGPs achieves best-in-class speed and scalability by limiting inference to the information contained in k nearest neighborhoods for prediction locations for both hyperparameter optimization and tuning. This feature affords the optimization of hyperparameters by way of leave-one-out cross-validation, as opposed to the more expensive loglikelihood evaluations requires by similar sparse methods

Priest, BenjaminW.↗

SDA: a symbolic differential algebra package in C++

Truncated Power Series Algebra (TPSA), or Differential Algebra (DA), is a well-established tool in accelerator physics, commonly used for generating high-order maps of dynamic systems, as well as in symplectic tracking, normal form analysis, verified integration, optimization, and fast multipole methods. This package is the first to perform symbolic DA computations, enabling traceability of initial condition contributions and runtime reduction for repeated DA calculations, potentially expanding DA’s applications.

97 MATHEMATICS AND COMPUTING↗

STARTR: An Open-Source MARVEL model for the NRIC Virtual Test Bed [Poster]

The National Reactor Innovation Center (NRIC) seeks to improve the understanding of microreactor physics in industry and academia through the development of a Microreactor Applications Research Validation and Evaluation (MARVEL) reactor-based model, published on the Virtual Test Bed (VTB). To achieve this goal, the Sodium-cooled Thermal-spectrum Advanced Research Test Reactor (STARTR) model was built using publicly available MARVEL specifications where possible and approximations where applicable, and was optimized for fast runtimes for researchers to receive rapid simulation feedback. STARTR will fill a gap between stakeholder interest and available models, as the first Sodium-cooled Thermal Reactor (STR) hosted on the VTB with baseline performance sanctioned by INL. This project involved the definition of all materials used in the reactor, geometry and all reactor subcomponents, and assertion of tallies and simulation settings within OpenMC 0.13.3. This poster details a small subset of the overall reactor physics testing: the two-dimensional power peaking factors and the flux energy spectrum, as well as plots of the created geometry. Future work includes code-to-code verification between the OpenMC-based model and a separately designed MCNP 6.2-based model.

21 - SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLAN↗

Power and Limitations of Linear Programming Decoder for Quantum LDPC Codes

Decoding quantum error-correcting codes is a key challenge in enabling fault-tolerant quantum computation. In the classical setting, linear programming (LP) decoders offer provable performance guarantees and can leverage fast practical optimization algorithms. Although LP decoders have been proposed for quantum codes, their performance and limitations remain relatively underexplored. In this work, we uncover a key limitation of LP decoding for quantum low-density parity-check (LDPC) codes: certain constant-weight error patterns lead to ambiguous fractional solutions that cannot be resolved through independent rounding. To address this issue, we incorporate a post-processing technique known as ordered statistics decoding (OSD), which significantly enhances LP decoding performance in practice. Our results show that LP decoding, when augmented with OSD, can outperform belief propagation with the same post-processing for intermediate code sizes of up to hundreds of qubits. These findings suggest that LP-based decoders, equipped with effective post-processing, offer a promising approach for decoding near-term quantum LDPC codes.

Gu, Shouzhen [Yale U.]↗

Power and Limitations of Linear Programming Decoder for Quantum LDPC Codes

Decoding quantum error-correcting codes is a key challenge in enabling fault-tolerant quantum computation. In the classical setting, linear programming (LP) decoders offer provable performance guarantees and can leverage fast practical optimization algorithms. Although LP decoders have been proposed for quantum codes, their performance and limitations remain relatively underexplored. In this work, we uncover a key limitation of LP decoding for quantum low-density parity-check (LDPC) codes: certain constant-weight error patterns lead to ambiguous fractional solutions that cannot be resolved through independent rounding. To address this issue, we incorporate a post-processing technique known as ordered statistics decoding (OSD), which significantly enhances LP decoding performance in practice. Our results show that LP decoding, when augmented with OSD, can outperform belief propagation with the same post-processing for intermediate code sizes of up to hundreds of qubits. These findings suggest that LP-based decoders, equipped with effective post-processing, offer a promising approach for decoding near-term quantum LDPC codes.

Gu, Shouzhen [Yale U.]↗

Power and Limitations of Linear Programming Decoder for Quantum LDPC Codes

Decoding quantum error-correcting codes is a key challenge in enabling fault-tolerant quantum computation. In the classical setting, linear programming (LP) decoders offer provable performance guarantees and can leverage fast practical optimization algorithms. Although LP decoders have been proposed for quantum codes, their performance and limitations remain relatively underexplored. In this work, we uncover a key limitation of LP decoding for quantum low-density parity-check (LDPC) codes: certain constant-weight error patterns lead to ambiguous fractional solutions that cannot be resolved through independent rounding. To address this issue, we incorporate a post-processing technique known as ordered statistics decoding (OSD), which significantly enhances LP decoding performance in practice. Our results show that LP decoding, when augmented with OSD, can outperform belief propagation with the same post-processing for intermediate code sizes of up to hundreds of qubits. These findings suggest that LP-based decoders, equipped with effective post-processing, offer a promising approach for decoding near-term quantum LDPC codes.

Gu, Shouzhen [Yale U.]↗

Performance-Aligned LLMs for Generating Fast HPC Code

Optimizing scientific software is a difficult task because codebases are often large and complex, and performance can depend upon several factors including the algorithm, its implementation, and hardware among others. Causes of poor performance can originate from disparate sources and be difficult to diagnose. Recent years have seen a multitude of work that use large language models (LLMs) to assist in software development tasks. However, these tools are trained to model the distribution of code as text, and are not specifically designed to understand performance aspects of code. In this work, we introduce a reinforcement learning based methodology to align the outputs of code LLMs with performance. This allows us to build upon the current code modeling capabilities of LLMs and extend them to generate better performing code. Here, we demonstrate that our fine-tuned model improves the expected speedup of generated code over base models for a set of benchmark tasks from 0.9 to 1.6 for serial code and 1.9 to 4.5 for OpenMP parallel code.

Computer science↗

Measuring Plant Metabolite Abundance in Spearmint ( Mentha spicata L.) with Raman Spectra to Determine Optimal Harvest Time

A fast field-deployable method utilizing Raman spectroscopy to determine the optimal harvest time of plants to extract the highest abundance of target metabolites is presented. Rosmarinic acid is a metabolite extracted from spearmint (Mentha spicata L.). Leaves from commercial “Native” and proprietary clonal line “KI110” spearmint were measured as a function of cell type and age to determine rosmarinic acid abundance. A linear regression model with leave-one-out cross-validation (R 2 CV = 0.61, RMSECV = 11.1 mg/g) was developed between selected Raman peak areas and rosmarinic acid concentrations determined by high-performance liquid chromatography (HPLC). A principal component analysis (PCA) model was also developed to determine rosmarinic acid abundance. The method may be suited to the analysis of many agriculturally relevant plant species and metabolites with distinct Raman peaks.

59 BASIC BIOLOGICAL SCIENCES↗

Carbon-Binder Optimization for Lithium-Ion Battery Extreme Fast Charge

Battery performance is strongly correlated with electrode microstructure and weight loading of the electrode components. Among them are the carbon-black and binder additives that enhance effective conductivity and provide mechanical integrity. However, these both reduce effective ionic transport in the electrolyte phase and reduce energy density. Therefore, an optimal additive loading is required to maximize performance, especially for fast charging where ionic transport is essential. Such optimization analysis is however challenging due to the nanoscale imaging limitations that prevent characterizing this additive phase and thus quantifying its impact on performance. Herein, an additive-phase generation algorithm has been developed to remedy this limitation and identify percolation threshold used to define a minimal additive loading. Improved ionic transport coefficients from reducing additive loading has been then quantified through homogenization calculation, macroscale model fitting, and experimental symmetric cell measurement, with good agreement between the methods. Rate capability test demonstrates capacity improvement at fast charge at the beginning of life, from 37% to 55%, respectively for high and low additive loading during 6C CC charging, in agreement with macroscale model, and attributed to a combination of lower cathode impedance, reduced electrode tortuosity and cathode thickness.

carbon-binder additives↗

Economic Storage Size Optimization for Electric Vehicle Extreme-Fast Charging Stations

En-route charging infrastructure for electric vehicles is critical to support transportation needs. These charging stations are likely to have high loads and especially sharp peak loads given fast charging capabilities needed to meet transportation schedules. In order to reduce both strain on distribution grid infrastructure and charging station operational costs, many stations are likely to employ behind the meter storage. This paper demonstrates a behind the meter storage sizing optimization that employs an open-source agent-based vehicle behavior model (BEAM) to determine the best sizing across many scenarios. This optimization and analysis is novel in that it examines how storage size impacts not only charging station cost and peak load, but also vehicle queue times. The optimization is also applied across a wide analysis region with sufficient diversity and numbers to provide novel statistical analysis of optimal sizes.

Aka, Julius↗

Fast multiscale contrast independent preconditioners for linear elastic topology optimization problems

The goal of this work is to present a fast and viable approach for the numerical solution of the high-contrast state problems arising in topology optimization. The optimization process is iterative, and the gradients are obtained by an adjoint analysis, which requires the numerical solution of large high-contrast linear elastic problems with features spanning several length scales. The size of the discretized problems forces the utilization of iterative linear solvers with solution time dependent on the quality of the preconditioner. The lack of clear separation between the scales, as well as the high-contrast, imposes severe challenges on the standard preconditioning techniques. Thus, here we propose new methods for the high-contrast elasticity equation with performance independent of the high-contrast and the multi-scale structure of the elasticity problem. The solvers are based on two-levels domain decomposition techniques with a carefully constructed coarse level to deal with the high-contrast and multi-scale nature of the problem. The construction utilizes spectral equivalence between scalar diffusion and each displacement block of the elasticity problems and, in contrast to previous solutions proposed in the literature, is able to select the appropriate dimension of the coarse space automatically. The new methods inherit the advantages of domain decomposition techniques, such as easy parallelization and scalability. Finally, the presented numerical experiments demonstrate the excellent performance of the proposed methods.

97 MATHEMATICS AND COMPUTING↗