Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Resource Size”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

Queue wait time prediction in high performance computing (HPC) systems

High Performance Computing (HPC) systems are critical enablers for groundbreaking scientific research across various domains. Efficient resource allocation, facilitated by job scheduling, is paramount for maximizing the utilization of HPC systems. However, the variability in wait times for queued jobs poses challenges for users, necessitating accurate job wait time estimation. This paper explores the influence of job characteristics, including job size (the number of nodes requested and walltime), the queue to which the job is submitted and other resource requirements, on job wait times in leadership-class HPC systems. Focusing on the Theta Cray XC40 and Polaris machines at Argonne National Laboratory, the study evaluates the performance of different supervised learning algorithms in predicting job wait times. It also evaluates the impact of data preprocessing, including outlier detection, Principal Component Analysis (PCA), and feature selection, on the performance of wait time prediction models. The findings reveal insights into the relationship between job characteristics and wait times, offering a foundation for optimizing resource allocation and enhancing user experience. The methodologies and tools developed in this study are adaptable to other leadership-class HPC systems, providing a valuable contribution to the broader HPC community aiming to improve job scheduling efficiency and user satisfaction.

Okafor, Nwamaka↗

Optimizing design and dispatch of a resilient renewable energy microgrid for a South African hospital

Lack of access to reliable energy is a major concern for countries in sub-Saharan Africa. The national grids are unable to consistently satisfy demand. Therefore, users turn to distributed generation systems in the form of back-up generators. However, such systems are usually designed based on a rule of thumb. We employ a mixed-integer linear programming model that considers several options such as renewable energy, combined heat and power, and storage technologies, in addition to those on-site, to provide optimal design and dispatch decisions that minimize total cost. We apply this model to a case study for a hospital in South Africa, considering its need for reliable electricity in light of multiple outages that might occur over the course of a year, as well as its high heating and cooling loads. Our results show that optimal design and dispatch decisions for the distributed generation system address reliability challenges, regardless of the time at which they occur. And, these solutions yield millions of dollars in savings, suggesting that technologies such as the absorption chiller may be overlooked in typical designs; its integration can reduce demand charges even in the absence of combined heat and power. We show that total cost is most sensitive to changes in site electrical demand, followed by capital cost, fuel cost, photovoltaic production, and monthly demand charges; changes in fuel cost primarily affect system sizes of combined heat and power and the absorption chiller, while photovoltaic system size is more sensitive to the changes in capital and fuel costs, photovoltaic resource availability, and hourly electrical demand. Finally, an outage simulator demonstrates the ability of our optimized system to sustain with no interruptions in power five-hour outages with probability 1.0 and ten-hour outages with probability 0.65, significant improvements over 0.5 and 0.0, respectively, under a business-as-usual case.

24 POWER TRANSMISSION AND DISTRIBUTION↗

A fast and robust computational modeling approach for density and shape predictions in powder metallurgy hot isostatic pressing

Powder metallurgy hot isostatic pressing (PM-HIP) is an advanced manufacturing process that produces near-net-shape parts with high material utilization and uniform microstructures. PM-HIP is frequently used for producing small-scale parts with complicated geometries and is potentially economical for producing large-scale parts. However, excessive post-HIP shape distortions can reduce its effectiveness and economic advantage, especially for larger parts. A PM-HIP computational model can predict and help mitigate these distortions. However, due to complex deformation mechanisms and thermo-mechanical coupling present in PM-HIP processes, these non-linear computational models sometimes become numerically unstable. The numerical instabilities in these models can lead to very slow convergence or no convergence at all, which often translates to slow and unreliable models. These limitations are more pronounced in large models with complicated geometries. Hence, in this work, an alternative modeling approach is presented that improves numerical stability and computational performance. The presented approach achieves these improvements through approximating the fully coupled thermo-mechanical PM-HIP model as a decoupled model and adding inertial damping to the model’s mechanical part. In conclusion, a comparison with the fully coupled model indicated a slight dip in prediction accuracy (<5% error) but significant improvements in numerical stability (>20 times larger time step size) and computational performance (5-10 times speed-up with less computational resource usage) when using the presented approach.

Hot isostatic pressing↗

Leveraging unlabeled SEM datasets with self-supervised learning for enhanced particle segmentation

Scanning Electron Microscopes (SEMs) are widely used in experimental science laboratories, often requiring cumbersome and repetitive user analysis. Automating SEM image analysis processes is highly desirable to address this challenge. In particle sample analysis, Machine Learning (ML) has emerged as the most effective approach for particle segmentation. However, the time-intensive process of manually annotating thousands of SEM images limits the applicability of supervised learning approaches. Self-Supervised Learning (SSL) offers a promising alternative by enabling knowledge extraction from raw, unlabeled data. This study presents a framework for evaluating SSL techniques in SEM image analysis, focusing on novel methods leveraging the ConvNeXtV2 architecture for particle detection. A dataset comprising 25,000 SEM images is curated to benchmark these proposed SSL methods. The results demonstrate that ConvNeXtV2 models, with varying parameter counts, consistently outperform other techniques in particle detection across different length scales, achieving up to a 34% reduction in relative error compared to established SSL methods. Furthermore, an ablation study explores the relationship between dataset size and SSL performance, providing actionable insights for practitioners regarding model selection and resource efficiency. This research advances the integration of SSL into autonomous analysis pipelines and supports its application in accelerating materials science discovery.

Rettenberger, Luca↗

Acceleration of Power System Dynamic Simulations Using a Deep Equilibrium Layer and Neural ODE Surrogate

The dominant paradigm for power system dynamic simulation is to build system-level simulations by combining physics-based models of individual components. The sheer size of the system along with the rapid integration of inverter-based resources exacerbates the computational burden of running time domain simulations. Here, in this paper, we propose a data-driven surrogate model based on implicit machine learningspecifically deep equilibrium layers and neural ordinary differential equationsto learn a reduced order model of a portion of the full underlying system. The data-driven surrogate achieves similar accuracy and reduction in simulation time compared to a physics-based surrogate, without the constraint of requiring detailed knowledge of the underlying dynamic models. This work also establishes key requirements needed to integrate the surrogate into existing simulation workflows; the proposed surrogate is initialized to a steady state operating point that matches the power flow solution by design.

Neural ordinary differential equations↗

Optimal Electrification Using Renewable Energies: Microgrid Installation Model with Combined Mixture k-Means Clustering Algorithm, Mixed Integer Linear Programming, and Onsset Method

Optimal planning and design of microgrids are priorities in the electrification of off-grid areas. Indeed, in one of the Sustainable Development Goals (SDG 7), the UN recommends universal access to electricity for all at the lowest cost. Several optimization methods with different strategies have been proposed in the literature as ways to achieve this goal. This paper proposes a microgrid installation and planning model based on a combination of several techniques. The programming language Python 3.10 was used in conjunction with machine learning techniques such as unsupervised learning based on K-means clustering and deterministic optimization methods based on mixed linear programming. These methods were complemented by the open-source spatial method for optimal electrification planning: onsset. Four levels of study were carried out. The first level consisted of simulating the model obtained with a cluster, which is considered based on the elbow and k-means clustering method as a case study. The second level involved sizing the microgrid with a capacity of 40 kW and optimizing all the resources available on site. The example of the different resources in the Togo case was considered. At the third level, the work consisted of proposing an optimal connection model for the microgrid based on voltage stability constraints and considering, above all, the capacity limit of the source substation. Finally, the fourth level involved a planning study of electrification strategies based mainly on microgrids according to the study scenario. The results of the first level of study enabled us to obtain an optimal location for the centroid of the cluster under consideration, according to the different load positions of this cluster. Then, the results of the second level of study were used to highlight the optimal resources obtained and proposed by the optimization model formulated based on the various technology costs, such as investment, maintenance, and operating costs, which were based on the technical limits of the various technologies. In these results, solar systems account for 80% of the maximum load considered, compared to 7.5% for wind systems and 12.5% for battery systems. Next, an optimal microgrid connection model was proposed based on the constraints of a voltage stability limit estimated to be 10% of the maximum voltage drop. The results obtained for the third level of study enabled us to present selective results for load nodes in relation to the source station node. Finally, the last results made it possible to plan electrification using different network technologies and systems in the short and long term. The case study of Togo was taken into account. The various results obtained from the different techniques provide the necessary leads for a feasibility study for optimal electrification of off-grid areas using microgrid systems.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Energy Transitions Initiative Partnership Project: City and Borough of Sitka, Alaska - Modeling and Controls Assistance and Renewable Energy Resource Assessment [Slides]

This presentation provides a summary of the ETIPP project objectives and findings for Sitka, Alaska, including sizing of wind penetration, dynamic models, and analysis of efficiency of load control, stability and grid control impacts of wind capacity expansions and locations, and wind-hydro control coordination.

17 WIND ENERGY↗

Government and industry interactions in the development of clock technology

It appears likely that everyone in the time and frequency community can agree on goals to be realized through the expenditure of resources. These goals are the same as found in most fields of technology: lower cost, better performance, increased reliability, small size and lower power. Related aspects are examined in the process of clock and frequency standard development. Government and industry are reviewed in a highly interactive role. These interactions include judgements on clock performance, what kind of clock, expenditure of resources, transfer of ideas or hardware concepts from government to industry, and control of production. Successful clock development and production requires a government/industry relationship which is characterized by long-term continuity, multidisciplinary team work, focused funding and a separation of reliability and production oriented tasks from performance improvement/research type efforts.

Hellwig, H.↗

Adversarial super-resolution of climatological wind and solar data

Accurate and high-resolution data reflecting different climate scenarios are vital for policy makers when deciding on the development of future energy resources, electrical infrastructure, transportation networks, agriculture, and many other societally important systems. However, state-of-the-art long-term global climate simulations are unable to resolve the spatiotemporal characteristics necessary for resource assessment or operational planning. We introduce an adversarial deep learning approach to super resolve wind velocity and solar irradiance outputs from global climate models to scales sufficient for renewable energy resource assessment. Using adversarial training to improve the physical and perceptual performance of our networks, we demonstrate up to a 50 × resolution enhancement of wind and solar data. In validation studies, the inferred fields are robust to input noise, possess the correct small-scale properties of atmospheric turbulent flow and solar irradiance, and retain consistency at large scales with coarse data. An additional advantage of our fully convolutional architecture is that it allows for training on small domains and evaluation on arbitrarily-sized inputs, including global scale. We conclude with a super-resolution study of renewable energy resources based on climate scenario data from the Intergovernmental Panel on Climate Change’s Fifth Assessment Report.

14 SOLAR ENERGY↗

Real-time semantic segmentation on FPGAs for autonomous vehicles with hls4ml

In this paper, we investigate how field programmable gate arrays can serve as hardware accelerators for real-time semantic segmentation tasks relevant for autonomous driving. Considering compressed versions of the ENet convolutional neural network architecture, we demonstrate a fully-on-chip deployment with a latency of 4.9 ms per image, using less than 30% of the available resources on a Xilinx ZCU102 evaluation board. The latency is reduced to 3 ms per image when increasing the batch size to ten, corresponding to the use case where the autonomous vehicle receives inputs from multiple cameras simultaneously. We show, through aggressive filter reduction and heterogeneous quantization-aware training, and an optimized implementation of convolutional layers, that the power consumption and resource utilization can be significantly reduced while maintaining accuracy on the Cityscapes dataset.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Enabling machine learning-ready HPC ensembles with Merlin

With the growing complexity of computational and experimental facilities, many scientific researchers are turning to machine learning (ML) techniques to analyze large scale ensemble data. With complexities such as multi-component workflows, heterogeneous machine architectures, parallel file systems, and batch scheduling, care must be taken to facilitate this analysis in a high performance computing (HPC) environment. Here, we present Merlin, a workflow framework to enable large ML-friendly ensembles of scientific HPC simulations. By augmenting traditional HPC with distributed compute technologies, Merlin aims to lower the barrier for scientific subject matter experts to incorporate ML into their analysis. As a producer–consumer workflow model, Merlin enables multi-machine, cross-batch job, dynamically allocated yet persistent workflows capable of utilizing surge-compute resources. Key features of Merlin are a flexible HPC-centric interface, low per-task overhead, multi-tiered fault recovery, and a hierarchical sampling algorithm that allows for $\mathscr{O}$(N) task execution and $\mathscr{O}$(N ln N) task queuing to ensembles of millions of tasks. In addition to Merlin’s design, we test the algorithm’s performance in an HPC center and demonstrate the ability to enqueue 40 million simulations in 100 s, with a 30 millisecond per-task overhead that is independent of ensemble size. Finally, we describe some example applications that Merlin has enabled on leadership-class HPC resources, such as the ML-augmented optimization of nuclear fusion experiments and the calibration of infectious disease models to study the progression of and possible mitigation strategies for COVID-19.

97 MATHEMATICS AND COMPUTING↗

Northern Colorado Water Resources: Using Earth Observations to Assess Relationships between Snowpack and Wildfires for Water Utility Management

In recent years, wildfires in the western U.S. have increased in size, frequency, and severity. Colorado is no exception to this trend, with many recent wildfires raising concerns. The Cameron Peak and East Troublesome Fires, the two largest recorded wildfires in Colorado history, burned over 400,000 acres in 2020. Wildfires pose a significant threat to high-elevation ecosystems, especially in terms of disturbances to watershed resources. Colorado’s annual water supply is dependent upon the melt and runoff of high-elevation snowpack, thus understanding wildfire effects is crucial. This project partnered with Northern Water, a water management company in Colorado, to explore fire impacts on Colorado Front Range watersheds and water supply forecasts. We investigated the use of remote sensing to study the influences of wildfire on snow depth, as Northern Water had previously relied on field-based observations of snowpack for their water supply forecasts. This project utilized airborne LiDAR data from Airborne Snow Observatories, Inc., Landsat 8 OLI derived products, as well as in situ snow survey data from Colorado State University to assess and model changes in snowpack over time. We uncovered key drivers of change in snowpack such as burn severity and snow zone and created a random forest model to determine and investigate their relationships. Despite temporal data availability limitations, analysis also showed that remotely-sensed snow depth measurements were strongly correlated with in situ snow depth measurements, revealing the accuracy and feasibility of using remote sensing technologies to study landscape-scale snowpack characteristics for water utility management purposes.

airborne LiDAR↗

A Scoping Study to Determine the Location-Specific WEC Threshold Size for Wave-Powered AUV Recharging

The aim of this study is to determine the threshold wave energy converter (WEC) type and size to charge a fleet of U.S. Navy autonomous underwater vehicles (AUVs) in various geographic locations of interest. The U.S. Navy deploys AUVs in locations around the world that must be charged manually, decreasing their operational endurance and creating operational limitations. Ocean waves are a potential power source that can be converted into electricity using a WEC and stored using a battery. It would be beneficial to develop a WEC that could autonomously charge AUVs offshore. Numerous locations were analyzed to determine the minimum size of a WEC capable of providing sufficient charging power and offering a strategic advantage. By predicting the WEC efficiency (based on empirical equations) and wave resource (based on available data), electrical power generation across numerous WEC types and locations was compared in MATLAB. The generalized process developed here could be used to determine the required size and type of WECs to charge a fleet of AUVs in different locations around the world.

16 TIDAL AND WAVE POWER↗

Visualizing and analyzing 3D biomolecular structures using Mol* at RCSB.org: Influenza A H5N1 virus proteome case study

The easiest and often most useful way to work with experimentally determined or computationally predicted structures of biomolecules is by viewing their three-dimensional (3D) shapes using a molecular visualization tool. Mol* was collaboratively developed by RCSB Protein Data Bank (RCSB PDB, RCSB.org) and Protein Data Bank in Europe (PDBe, PDBe.org) as an open-source, web-based, 3D visualization software suite for examination and analyses of biostructures. It is capable of displaying atomic coordinates and related experimental data of biomolecular structures together with a variety of annotations, facilitating basic and applied research, training, education, and information dissemination. Across RCSB.org, the RCSB PDB research-focused web portal, Mol* has been implemented to support single-mouse-click atomic-level visualization of biomolecules (e.g., proteins, nucleic acids, carbohydrates) with bound cofactors, small-molecule ligands, ions, water molecules, or other macromolecules. RCSB.org Mol* can seamlessly display 3D structures from various sources, allowing structure interrogation, superimposition, and comparison. Using influenza A H5N1 virus as a topical case study of an important pathogen, we exemplify how Mol* has been embedded within various RCSB.org tools—allowing users to view polymer sequence and structure-based annotations integrated from trusted bioinformatics data resources, assess patterns and trends in groups of structures, and view structures of any size and compositional complexity. In addition to being linked to every experimentally determined biostructure and Computed Structure Model made available at RCSB.org, Standalone Mol* is freely available for visualizing any atomic-level or multi-scale biostructure at rcsb.org/3d-view.

3D biostructure↗

Wind power costs driven by innovation and experience with further reductions on the horizon

The costs of wind power have declined to levels on par with or below those of conventional sources in many parts of the world. Wind power has become one of the fastest-growing sources of new electricity generation. We take stock of wind power cost evolution over the past 20 years, review methodologies commonly used for cost assessment, discuss the potential for continued cost reduction, and identify anticipated cost and value drivers. Our scope includes both onshore and offshore wind technologies. We draw from a vast body of literature on these topics to highlight key trends, approaches, and limitations. Furthermore, we discuss strategies for wind power assets to enhance their marginal economic value to the broader power system and consumers. We identify a myriad of factors that are expected to influence the future cost and value of wind power, including siting, project scale, turbine size, operational synergies, commodity prices, advancements in turbine technologies, enhanced management of the wind resource, and novel control technologies that provide value for the electricity grid. Because the common methods for forecasting future costs each have their own strengths and weaknesses, we find the best insights are elicited from a combination of methods. Overall, researchers and analysts anticipate further sizable cost reductions for onshore and offshore wind. Midrange forecasts for levelized cost of energy in 2050 are generally between $20 and $30/MWh for onshore wind and $40 and $60/MWh for offshore wind, a reduction to approximately half of today's levels. Optimistic forecasts anticipate these levels as early as 2030.

17 WIND ENERGY↗

Design and implementation of I/O performance prediction scheme on HPC systems through large-scale log analysis

Abstract Large-scale high performance computing (HPC) systems typically consist of many thousands of CPUs and storage units used by hundreds to thousands of users simultaneously. Applications from large numbers of users have diverse characteristics, such as varying computation, communication, memory, and I/O intensity. A good understanding of the performance characteristics of each user application is important for job scheduling and resource provisioning. Among these performance characteristics, I/O performance is becoming increasingly important as data sizes rapidly increase and large-scale applications, such as simulation and model training, are widely adopted. However, predicting I/O performance is difficult because I/O systems are shared among all users and involve many layers of software and hardware stack, including the application, network interconnect, operating system, file system, and storage devices. Furthermore, updates to these layers and changes in system management policy can significantly alter the I/O behavior of applications and the entire system. To improve the prediction of the I/O performance on HPC systems, we propose integrating information from several different system logs and developing a regression-based approach to predict the I/O performance. Our proposed scheme can dynamically select the most relevant features from the log entries using various feature selection algorithms and scoring functions, and can automatically select the regression algorithm with the best accuracy for the prediction task. The evaluation results show that our proposed scheme can predict the write performance with up to 90% prediction accuracy and the read performance with up to 99% prediction accuracy using the real logs from the Cori supercomputer system at NERSC.

97 MATHEMATICS AND COMPUTING↗

LANDSAT D instrument module study

Spacecraft instrument module configurations which support an earth resource data gathering mission using a thematic mapper sensor were examined. The differences in size of these two experiments necessitated the development of two different spacecraft configurations. Following the selection of the best-suited configurations, a validation phase of design, analysis and modelling was conducted to verify feasibility. The chosen designs were then used to formulate definition for a systems weight, a cost range for fabrication and interface requirements for the thematic mapper (TM).

Source record↗

Rdesign: A data dictionary with relational database design capabilities in Ada

Data Dictionary is defined to be the set of all data attributes, which describe data objects in terms of their intrinsic attributes, such as name, type, size, format and definition. It is recognized as the data base for the Information Resource Management, to facilitate understanding and communication about the relationship between systems applications and systems data usage and to help assist in achieving data independence by permitting systems applications to access data knowledge of the location or storage characteristics of the data in the system. A research and development effort to use Ada has produced a data dictionary with data base design capabilities. This project supports data specification and analysis and offers a choice of the relational, network, and hierarchical model for logical data based design. It provides a highly integrated set of analysis and design transformation tools which range from templates for data element definition, spreadsheet for defining functional dependencies, normalization, to logical design generator.

Lekkos, Anthony A.↗