Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “pattern classification”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

2013 Metropolitan Area Planning Agency External Travel Survey

# 2013 Metropolitan Area Planning Agency External Travel Survey The 2013 Metropolitan Area Planning Agency (MAPA) External Travel Survey was conducted to measure and identify travel patterns into, within, and out of the greater Omaha/Council Bluffs metropolitan area in Nebraska. MAPA sponsored the survey in conjunction with the Federal Highway Administration, the Nebraska Department of Roads, and the Iowa Department of Transportation. ## Data Collection Agency MAPA conducted the survey. ## Methodology The purpose of the survey was to collect information and data needed as input for MAPA’s travel-demand model. The survey employed a combination of nine survey methods and data-collection activities, including Bluetooth technology, intercept surveys, postcard handouts, travel-time studies, vehicle classification counts, and a web-based survey. ## Survey Records Survey records include a total of 729 participants. ## Transportation Data This study provides Bluetooth records and supplementary data for 17,434 passenger trips and 3,123 commercial trips, accounting for 714,218 vehicle miles traveled. Transportation data are available as zipped files. [Download Winzip](http://www.winzip.com/downwz.htm).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

2013 Metropolitan Area Planning Agency External Travel Survey

# 2013 Metropolitan Area Planning Agency External Travel Survey The 2013 Metropolitan Area Planning Agency (MAPA) External Travel Survey was conducted to measure and identify travel patterns into, within, and out of the greater Omaha/Council Bluffs metropolitan area in Nebraska. MAPA sponsored the survey in conjunction with the Federal Highway Administration, the Nebraska Department of Roads, and the Iowa Department of Transportation. ## Data Collection Agency MAPA conducted the survey. ## Methodology The purpose of the survey was to collect information and data needed as input for MAPA’s travel-demand model. The survey employed a combination of nine survey methods and data-collection activities, including Bluetooth technology, intercept surveys, postcard handouts, travel-time studies, vehicle classification counts, and a web-based survey. ## Survey Records Survey records include a total of 729 participants. ## Transportation Data This study provides Bluetooth records and supplementary data for 17,434 passenger trips and 3,123 commercial trips, accounting for 714,218 vehicle miles traveled. Transportation data are available as zipped files. [Download Winzip](http://www.winzip.com/downwz.htm).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

2013 Metropolitan Area Planning Agency External Travel Survey

# 2013 Metropolitan Area Planning Agency External Travel Survey The 2013 Metropolitan Area Planning Agency (MAPA) External Travel Survey was conducted to measure and identify travel patterns into, within, and out of the greater Omaha/Council Bluffs metropolitan area in Nebraska. MAPA sponsored the survey in conjunction with the Federal Highway Administration, the Nebraska Department of Roads, and the Iowa Department of Transportation. ## Data Collection Agency MAPA conducted the survey. ## Methodology The purpose of the survey was to collect information and data needed as input for MAPA’s travel-demand model. The survey employed a combination of nine survey methods and data-collection activities, including Bluetooth technology, intercept surveys, postcard handouts, travel-time studies, vehicle classification counts, and a web-based survey. ## Survey Records Survey records include a total of 729 participants. ## Transportation Data This study provides Bluetooth records and supplementary data for 17,434 passenger trips and 3,123 commercial trips, accounting for 714,218 vehicle miles traveled. Transportation data are available as zipped files. [Download Winzip](http://www.winzip.com/downwz.htm).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

2013 Metropolitan Area Planning Agency External Travel Survey

# 2013 Metropolitan Area Planning Agency External Travel Survey The 2013 Metropolitan Area Planning Agency (MAPA) External Travel Survey was conducted to measure and identify travel patterns into, within, and out of the greater Omaha/Council Bluffs metropolitan area in Nebraska. MAPA sponsored the survey in conjunction with the Federal Highway Administration, the Nebraska Department of Roads, and the Iowa Department of Transportation. ## Data Collection Agency MAPA conducted the survey. ## Methodology The purpose of the survey was to collect information and data needed as input for MAPA’s travel-demand model. The survey employed a combination of nine survey methods and data-collection activities, including Bluetooth technology, intercept surveys, postcard handouts, travel-time studies, vehicle classification counts, and a web-based survey. ## Survey Records Survey records include a total of 729 participants. ## Transportation Data This study provides Bluetooth records and supplementary data for 17,434 passenger trips and 3,123 commercial trips, accounting for 714,218 vehicle miles traveled. Transportation data are available as zipped files. [Download Winzip](http://www.winzip.com/downwz.htm).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

2013 Metropolitan Area Planning Agency External Travel Survey

# 2013 Metropolitan Area Planning Agency External Travel Survey The 2013 Metropolitan Area Planning Agency (MAPA) External Travel Survey was conducted to measure and identify travel patterns into, within, and out of the greater Omaha/Council Bluffs metropolitan area in Nebraska. MAPA sponsored the survey in conjunction with the Federal Highway Administration, the Nebraska Department of Roads, and the Iowa Department of Transportation. ## Data Collection Agency MAPA conducted the survey. ## Methodology The purpose of the survey was to collect information and data needed as input for MAPA’s travel-demand model. The survey employed a combination of nine survey methods and data-collection activities, including Bluetooth technology, intercept surveys, postcard handouts, travel-time studies, vehicle classification counts, and a web-based survey. ## Survey Records Survey records include a total of 729 participants. ## Transportation Data This study provides Bluetooth records and supplementary data for 17,434 passenger trips and 3,123 commercial trips, accounting for 714,218 vehicle miles traveled. Transportation data are available as zipped files. [Download Winzip](http://www.winzip.com/downwz.htm).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

2013 Metropolitan Area Planning Agency External Travel Survey

# 2013 Metropolitan Area Planning Agency External Travel Survey The 2013 Metropolitan Area Planning Agency (MAPA) External Travel Survey was conducted to measure and identify travel patterns into, within, and out of the greater Omaha/Council Bluffs metropolitan area in Nebraska. MAPA sponsored the survey in conjunction with the Federal Highway Administration, the Nebraska Department of Roads, and the Iowa Department of Transportation. ## Data Collection Agency MAPA conducted the survey. ## Methodology The purpose of the survey was to collect information and data needed as input for MAPA’s travel-demand model. The survey employed a combination of nine survey methods and data-collection activities, including Bluetooth technology, intercept surveys, postcard handouts, travel-time studies, vehicle classification counts, and a web-based survey. ## Survey Records Survey records include a total of 729 participants. ## Transportation Data This study provides Bluetooth records and supplementary data for 17,434 passenger trips and 3,123 commercial trips, accounting for 714,218 vehicle miles traveled. Transportation data are available as zipped files. [Download Winzip](http://www.winzip.com/downwz.htm).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

2013 Metropolitan Area Planning Agency External Travel Survey

# 2013 Metropolitan Area Planning Agency External Travel Survey The 2013 Metropolitan Area Planning Agency (MAPA) External Travel Survey was conducted to measure and identify travel patterns into, within, and out of the greater Omaha/Council Bluffs metropolitan area in Nebraska. MAPA sponsored the survey in conjunction with the Federal Highway Administration, the Nebraska Department of Roads, and the Iowa Department of Transportation. ## Data Collection Agency MAPA conducted the survey. ## Methodology The purpose of the survey was to collect information and data needed as input for MAPA’s travel-demand model. The survey employed a combination of nine survey methods and data-collection activities, including Bluetooth technology, intercept surveys, postcard handouts, travel-time studies, vehicle classification counts, and a web-based survey. ## Survey Records Survey records include a total of 729 participants. ## Transportation Data This study provides Bluetooth records and supplementary data for 17,434 passenger trips and 3,123 commercial trips, accounting for 714,218 vehicle miles traveled. Transportation data are available as zipped files. [Download Winzip](http://www.winzip.com/downwz.htm).

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Global Geo-processed Data of Aquifer Properties by 0.5° Grid, Country and Water Basins

This repository of global hydrogeologic datasets contains aquifer properties on 0.5° scale, including depth to groundwater (Fan et al., 2013), aquifer thickness (de Graaf et al., 2015), WHYMap aquifer classes (Richts et al., 2011), porosity and permeability (Gleeson et al., 2014), digitized and geo-processed from their respective sources. Globally gridded aquifer properties could be used independently to estimate global groundwater availability or used as critical inputs to the superwell model to simulate groundwater extraction and provide estimates of pumped volumes and unit costs under user-specific scenarios. Key resources related to this data are: Niazi, H., Ferencz, S., Graham, N., Yoon, J., Wild, T., Hejazi, M., Watson, D., & Vernon, C. (2024; In-prep). Long-term Hydro-economic Assessment Tool for Evaluating Global Groundwater Cost and Supply: Superwell v1. Geoscientific Model Development. superwell model repository which uses this data to simulate groundwater extraction and provides estimates of the global extractable volumes and unit-costs ($/km3) of accessible groundwater production under user-specified extraction scenarios. Repository Overview Main output: aquifer_properties.csv contains all processed outputs, including aquifer properties like porosity, permeability, aquifer thickness, and depth to groundwater. shapefiles.zip: contains all digitized GIS databases and shapefile for all aquifer properties prep_inputs.R: R script that processes the shapefiles to produce the aquifer_properties file plot_inputs.R: R script for plotting the maps and conducting preliminary analysis on the available groundwater volume basin_to_country_mapping.csv, basin_country_region_mapping.csv and continent_county_mapping.csv provide the mapping between continents, 32 energy-economic macro regions, countries, and water basins for post-processing aquifer_properties.csv Maps: Each map visualizes the spatial distribution of one of the aquifer properties across the globe map_in_Porosity.png map_in_Permeability.png map_in_Aquifer_thickness.png map_in_Depth_to_water.png map_in_Grid_area_km.png map_in_WHYClass.png Sample inputs sample_inputs.py: this script samples inputs from the aquifer_properties dataset, ensuring the sampled and original inputs maintain the same distributions sampled_data_100.csv contains 100 sampled data points and sampled_data_100.png compares their distributions Dataset Overview The main outputs are consolidated in a comprehensive aquifer_properties.csv file and include the following fields: GridCellID: Unique identifier for each (roughly 0.5°) grid cell Continent: Continent name Country: Country name GCAM_basin_ID: Identifier for GCAM hydrologic basin Basin_long_name: Full name of the basin WHYClass: Hydrogeologic classification based on WHYMap aquifer classes (Richts et al., 2011) Porosity: Soil porosity (%) (Gleeson et al., 2014) Permeability: Soil permeability (in square meters; Gleeson et al., 2014) Aquifer_thickness: Thickness of the aquifer (in meters; de Graaf et al., 2015) Depth_to_water: Depth to groundwater (in meters; Fan et al., 2013) Grid_area: Area of the grid cell (in square meters) Key References The datasets are digitized versions of global hydrogeologic properties from the following key literature sources: Depth to Groundwater: Fan, Y., Li, H., & Miguez-Macho, G. (2013). Global Patterns of Groundwater Table Depth. Science, 339(6122), 940-943. https://doi.org/10.1126/science.1229881 Aquifer Thickness: de Graaf, I. E. M., Sutanudjaja, E. H., van Beek, L. P. H., & Bierkens, M. F. P. (2015). A high-resolution global-scale groundwater model. Hydrol. Earth Syst. Sci., 19(2), 823-837. https://doi.org/10.5194/hess-19-823-2015 Porosity and Permeability: Gleeson, T., Moosdorf, N., Hartmann, J., & van Beek, L. P. H. (2014). A glimpse beneath earth's surface: GLobal HYdrogeology MaPS (GLHYMPS) of permeability and porosity. Geophysical Research Letters, 41(11), 3891-3898. https://doi.org/10.1002/2014GL059856 Aquifer classes: Richts, A., Struckmeier, W. F., & Zaepke, M. (2011). WHYMAP and the Groundwater Resources Map of the World 1:25,000,000. In J. A. A. Jones (Ed.), Sustaining Groundwater Resources: A Critical Element in the Global Water Crisis (pp. 159-173). Springer Netherlands. https://doi.org/10.1007/978-90-481-3426-7_10 Cite as Niazi, H., Watson, D., Hejazi, M., Yonkofski, C., Ferencz, S., Vernon, C., Graham, N., Wild, T., & Yoon, J. (2024). Global Geo-processed Data of Aquifer Properties by 0.5° Grid, Country and Water Basins. MSD-LIVE Data Repository. https://doi.org/10.57931/2307831 Contact Reach out to Hassan Niazi or open an issue in superwell repository for questions or suggestions.

aquifer thickness↗

A Taxonomic Classification Approach for Global Spatio-temporal Data

The World Bank, World Health Organization, and other major vendors collectively provide thousands of global time series datasets that focus on issues of the environment, public health, economics, violence, education, and national security. Sorting these data into meaningful information requires the use of data mining techniques to cluster trends into an orderly and manageable number of cases. The World SpatioTemporal Analytics and Mapping (WSTAMP) project database (wstamp.ornl.gov) was developed to spatiotemporally harmonize global vendor data (23,300+ attributes, 200+ locations, 50+ years). Within the WSTAMP analytical environment, Dynamic Time Warping (DTW) has been a highly effective data-driven approach for clustering and mapping these time series into national spatiotemporal behavior maps. Two significant properties have surfaced from this work. First, several recognizable cluster patterns have emerged and persist across a range of locations, attributes, and time frames (e.g., increasing, decreasing, rebounding, peak, oscillating). Secondly, practitioners engaging WSTAMP have noted the explanatory and anticipatory value of these patterns and articulated particular interest in detecting them within the spatiotemporal cube. This need was addressed by shifting DTW-based clustering from an open ended, data-driven implementation to a taxonomic pattern matching approach. This paper presents the method including implementation strategies for visualization and human computer interaction and applies the approach to a sample data set and concludes with next steps.

Stewart, Robert↗

Grass Evolutionary Lineages Can Be Identified Using Hyperspectral Leaf Reflectance

Hyperspectral remote sensing has the potential to map numerous attributes of the Earth’s surface, including spatial patterns of biological diversity. Grasslands are one of the largest biomes on Earth. Accurate mapping of grassland biodiversity relies on spectral discrimination of endmembers of species or plant functional types. We focused on spectral separation of grass lineages that dominate global grassy biomes: Andropogoneae (C 4 ), Chloridoideae (C 4 ), and Pooideae (C 3 ). We examined leaf reflectance spectra (350–2,500 nm) from 43 grass species representing these grass lineages from four representative grassland sites in the Great Plains region of North America. Here, we assessed the utility of leaf reflectance data for classification of grass species into three major lineages and by collection site. Classifications had very high accuracy (94%) that were robust to site differences in species and environment. We also show an information loss using multispectral sensors, that is, classification accuracy of grass lineages using spectral bands provided by current multispectral satellites is much lower (accuracy of 85.2% and 61.3% using Sentinel 2 and Landsat 8 bands, respectively). Our results suggest that hyperspectral data have an exciting potential for mapping grass functional types as informed by phylogeny. Leaf-level hyperspectral separability of grass lineages is consistent with the potential increase in biodiversity and functional information content from the next generation of satellite-based spectrometers.

59 BASIC BIOLOGICAL SCIENCES↗

A Mixed Length Scale Model for Migrating Fluvial Bedforms

With the expansion of hydropower, in-stream converters, flood-protection infrastructures, and growing concerns on deltas fragile ecosystems, there is a pressing need to evaluate and monitor bedform sediment mass flux. It is critical to estimate real-time bedform size and migration velocity and provide a theoretical framework to convert easily accessible time histories of bed elevations into spatially evolving patterns. In this study, we collected spatiotemporally resolved bathymetries from laboratory flumes and the Colorado River in statistically steady, homogeneous, subcritical flow conditions. Wave number and frequency spectra of bed elevations show compelling evidence of scale-dependent velocity for the hierarchy of migrating bedforms observed in the laboratory and field. New scaling laws were applied to describe the full range of migration velocities as function of two dimensionless groups based on the bed shear velocity, sediment diameter, and water depth. Further simplification resulted in a mixed length scale model estimating scale-dependent migration velocities, without requiring bedform classification or identification.

58 GEOSCIENCES↗

Phasor-Measurement-Unit-Based Data Analytics Using Digital Twin and PhasorAnalytics Software

A major objective of this project was to apply GE’s commercial machine learning and data analytics toolsets to large-scale, real-world, anonymized Phasor Measurement Unit (PMU) datasets in order to extract signatures, correlated and/or causal factors, and precursor patterns associated with significant power system phenomena. The project had a particular emphasis on extraction of insights relevant to asset health monitoring, real-time load modeling and cybersecurity monitoring. Additionally, the team was directed to undertake a comprehensive data quality analysis for the provided datasets and encouraged to estimate the ‘machine-learning readiness’ of the datasets by documenting any major obstacles to the application of commercial machine learning algorithms. To accomplish the aforementioned objectives, the project team’s work centered around the identification of key event signatures and application of the identified event signatures for event detection and event classification. The industry-validated, semi-supervised machine learning strategy employed for event signature identification involved several major tasks, including data-preprocessing, generation of an overabundance of features, normal data identification, normality modeling, and event signature identification through a methodical, quantitative ranking of features in order of relevance to each studied event type. Throughout the project, data quality issues and mitigation techniques were investigated. In this report, insights are provided regarding the readiness of the provided synchrophasor datasets for application of machine learning and data analytics. The methodologies employed for this technical strategy are summarized in this report. With regards to data preprocessing and feature generation, the provided Training and Test Datasets were ingested into GE’s big data environment. Subsequently, the team applied bad data cleansing and data imputation scripts, event detection scripts, and application programming interfaces (APIs) to the datasets for convenient data access. The project team completed development and validation of dozens of physics-based, statistics-based and transformation-based feature functions used for the extraction of over 60 synchrophasor features. Using a new parallel feature generation technology developed on this project, over 60 features have been rapidly generated for the full two years’ worth of Training and Test Dataset data associated with both the Eastern and Western interconnects. Even accommodating for temporal down-sampling inherent to the feature extraction procedure, this parallel feature generation activity resulted in a massive feature set with a storage requirement approximately equal to that of the raw training dataset itself. With regards to normal data identification and normality modeling, a normality model was built using the feature data extracted from the Training Dataset and iteratively refined subsequent to incremental adjustments and expansions of the Training Dataset feature data. With respect to event characterization and signature identification, an event signature identification pipeline was developed and used in conjunction with the normality model to identify over 15 event signatures for key event categories within the Training Dataset. The identified event signatures were used to characterize hundreds of key events in terms of relative severity, duration, and location of the event. An investigation was undertaken to identify correlated and causal factors involved in transformer events. A separate investigation into temporal trends in ring-down analysis results was undertaken to determine possible associations between system dynamics and various other factors such as loading, season or year. To validate the identified event signatures, additional work was undertaken to develop signature-based anomaly detection and classification tools suitable for convenient application to the synchrophasor datasets. The anomaly detection and classification tools, suitable for online application, were then applied to the entirety of the Eastern Interconnect Training and Test Datasets. Performance of the event detection and classification tools was evaluated upon receipt of the Test Dataset event logs (i.e., the labels for events contained in the Test Dataset), and promising results were obtained despite several challenges (documented herein) associated with application of supervised or semi-supervised machine learning methods to large-scale, anonymized datasets. Finally, the detection and classification tools were used to detect, classify, and characterize thousands of new events not included in the original event logs provided by the DOE within both the Training and Test Datasets.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Explainable Machine Learning for Functional Data

Black-box machine learning models are recognized as useful tools for prediction applications, but the algorithmic complexity of some models causes interpretation challenges. Explainability methods have been proposed to provide insight into these models, but there is little research focused on supervised modeling with functional data inputs. We argue that, especially in applications of high consequence, it is important to explicitly model the functional dependence in a black-box analysis to not obscure or misrepresent patterns in explanations. As such, we propose the V ariable importance E xplainable E lastic S hape A nalysis (VEESA) pipeline for training supervised machine learning models with functional inputs. The pipeline is an analysis process that includes the data preprocessing, modeling, and post-hoc explanations. The preprocessing is done using elastic functional principal components analysis, which accounts for vertical and horizontal variability in functional data and, ultimately, allows for explanations in the original data space that identify the important functional variability without bias due to correlated variables. Here, we demonstrate the pipeline on two high-consequence applications: explosives classification for national security and inkjet printer identification in forensic science. The applications exhibit the VEESA pipeline’s ability to provide an understanding of the characteristics of the functional data useful for prediction. Code for implementing the pipeline is available in the veesa R package (and supplemental python code).

Elastic Shape Analysis↗

Data-Driven Whole-Genome Clustering to Detect Geospatial, Temporal, and Functional Trends in SARS-CoV-2 Evolution

Current methods for defining SARS-CoV-2 lineages ignore the vast majority of the SARS-CoV-2 genome. We develop and apply an exhaustive vector comparison method that directly compares all known SARS-CoV-2 genome sequences to produce novel lineage classifications. We utilize data-driven models that (i) accurately capture the complex interactions across the set of all known SARS-CoV-2 genomes, (ii) scale to leadership-class computing systems, and (iii) enable tracking how such strains evolve geospatially over time. We show that during the height of the original Omicron surge, countries across Europe, Asia, and the Americas had a spatially asynchronous distribution of Omicron sub-strains. Moreover, neighboring countries were often dominated by either different clusters of the same variant or different variants altogether throughout the pandemic. Analyses of this kind may suggest a different pattern of epidemiological risk than was understood from conventional data, as well as produce actionable insights and transform our ability to prepare for and respond to current and future biological threats.

Jacobson, Daniel↗

Simulating water dynamics related to pedogenesis across space and time: Implications for four-dimensional digital soil mapping

Digital soil mapping (DSM) relies on machine-learning and geostatistics to represent soil property observations across space. DSM techniques are powerful but often empirical, being limited to the quality and density of point samples. Water dynamics are closely related to soil variability, and the physics that govern water movement are well known. Hydrological properties can hence be simulated by physical models through space and time, unveiling key characteristics about soils. We propose the use of hydrologic models to map soils across the surface (2D), depth (1D), and time (1D)–which provides a 4D approach to digital soil mapping (4DSM). The Distributed Hydrology Soil Vegetation Model (DHSVM) was applied to a watershed currently under pasture. Moisture sensors and wells were installed at different depths in the watershed on summit, sideslope and toeslope positions to validate the model. DHSVM simulations of soil moisture distribution and depth to saturation were performed during the hydrological year (October 2008-September 2009). Clusters of similar pixels based on soil moisture values were determined using Dynamic Time Warping (DTW) to align temporal data and K-means. Clustering was performed both seasonally and for the entire year. Temporal patterns simulated by DHSVM matched measurements given by moisture sensors and wells. Seasonal clusters differed from the annual cluster. Distinct clusters were observed for each season and with depth, showing that spatiotemporal soil variability is lost when statically assessing soils. Spatiotemporal clusters corroborated field observations of fragipan occurrence not explicitly spatially mapped by Soil Survey Geographic Database (SSURGO). If a connection can be made between water and soils, static and dynamic soil variability can be predicted using physically based hydrologic models. Hydrologic models can benefit soil mapping by enabling reliable 4D simulation of water dynamics, which are fundamental to soil variability and soil classification and directly relate to biological, physical and chemical soil processes not captured by typical soil sampling protocols.

54 ENVIRONMENTAL SCIENCES↗

A derecho climatology (2004–2021) in the United States based on machine learning identification of bow echoes

Due to their persistent widespread severe winds, derechos pose significant threats to human safety and property, with impacts comparable to many tornadoes and hurricanes. Yet, automated detection of derechos remains challenging due to the absence of spatiotemporally continuous observations and the complex criteria employed to define the phenomenon. This study presents an objective derecho detection approach capable of automatically identifying derechos through both observations and model results. The approach is grounded in a physically based definition of derechos and integrates three algorithms: (1) the Python Flexible Object Tracker (PyFLEXTRKR) algorithm to track mesoscale convective systems (MCSs), (2) a semantic segmentation convolutional neural network to identify bow echoes, and (3) a comprehensive classification algorithm to detect derechos within MCS life cycles and distinguish derecho-producing from non-derecho-producing MCSs. Using this approach, we developed a novel high-resolution (4 km and hourly) observational dataset of derechos and accompanying derecho-producing MCSs over the United States east of the Rocky Mountains from 2004 to 2021. The dataset consists of two subsets based on different gust speed data sources and is analyzed to document the climatology of derechos in the United States. On average, 12–15 derechos are identified per year, aligning with previous estimations (∼6–21 events annually). The spatial distribution and seasonal variation patterns are consistent with prior studies, showing peak occurrences in the Great Plains and the Midwest during the warm season. Additionally, during the study period, derechos account for approximately 3.1 % of measured damaging gusts (≥25.93 m s−1) over the eastern United States. The dataset is publicly available at https://doi.org/10.5281/zenodo.14835362 (Li et al., 2025).

54 ENVIRONMENTAL SCIENCES↗

Quantum machine learning for chemistry and physics

Machine learning (ML) has emerged as a formidable force for identifying hidden but pertinent patterns within a given data set with the objective of subsequent generation of automated predictive behavior. In recent years, it is safe to conclude that ML and its close cousin, deep learning (DL), have ushered in unprecedented developments in all areas of physical sciences, especially chemistry. Not only classical variants of ML, even those trainable on near-term quantum hardwares have been developed with promising outcomes. Such algorithms have revolutionized materials design and performance of photovoltaics, electronic structure calculations of ground and excited states of correlated matter, computation of force-fields and potential energy surfaces informing chemical reaction dynamics, reactivity inspired rational strategies of drug designing and even classification of phases of matter with accurate identification of emergent criticality. In this review we shall explicate a subset of such topics and delineate the contributions made by both classical and quantum computing enhanced machine learning algorithms over the past few years. We shall not only present a brief overview of the well-known techniques but also highlight their learning strategies using statistical physical insight. The objective of the review is not only to foster exposition of the aforesaid techniques but also to empower and promote cross-pollination among future research in all areas of chemistry which can benefit from ML and in turn can potentially accelerate the growth of such algorithms.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Design, Detection, and Countermeasure of Frequency Spectrum Attack and Its Impact on Long Short-Term Memory Load Forecasting and Microgrid Energy Management

This paper introduces a frequency-domain false data injection attack called Frequency Spectrum Attack (FSA) and explores its effects on load forecasting and the energy management system (EMS) in a microgrid. The FSA analyzes time-series signals in the frequency domain to identify patterns in their frequency spectrum. It learns the distribution of dominant frequencies in a dataset of healthy signals. Subsequently, it manipulates the amplitudes of dominant frequencies within this healthy distribution, ensuring a stealthy attack against statistical analysis of the signal spectrum. We evaluated the performance of FSA on LSTM, a state-of-the-art network for load forecasting. The results show that FSA can triple the Mean Absolute Error (MAE) of predictions compared to the normal case and increase it by 70% compared to noise injection attacks. Furthermore, FSA indirectly enhances battery utilization in the EMS by 45%. We then proposed a detection method that combines statistical analysis and machine-learning-based classification techniques with features. The model effectively distinguishes FSA from healthy and noisy signals, achieving an accuracy of 98.7% and an F1-score of 98.1% on a load dataset, covering healthy, FSA, and noisy load data. Finally, a countermeasure was introduced based on the statistical analysis of the frequency spectrum of healthy signals to mitigate the impact of FSA. This countermeasure successfully reduces the MAE of the attacked model from 0.135 to 0.053, validating its effectiveness in mitigating FSA.

Nazeri, Amirhossein↗