Engineering PapersSearch

SEARCH · Engineering Papers

Results for “population forecasting”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

PopGNN: Graph Neural Network-Based Flexible Future Population Forecasting Model

Accurate population forecasts is important to plan critical infrastructure and services, from housing and education to healthcare and transport. However, traditional population prediction studies have only employed traditional machine learning models limited to capture complex spatial interdependencies and patterns. Althogh recently computer vision-based framework was introduced with with promising accuracy, it has critical limitations for real-world planning applications: it function only at fixed spatial resolutions, restricting their use in diverse boundaries such as census tracts, neighborhoods, or administrative zones. Therefore, this study suggests a Graph Neural Network (GNN)-based population prediction framework, called PopGNN. This model recorded remarkable performance compared with state-of-the-art models and traditional baseline models in the grid and administrative boundaries. Furthermore, our framework achieved comparable predictive accuracy to a computer vision-based model in both the South Korea and Tennessee case studies. Consequently, this study is valuable in that a single model can provide accurate population forecasts that address diverse planning demands, ranging from granular grid-level estimates for precise service allocation and facility location planning to aggregate administrative-level forecasts for macro-scale regional policy and resource distribution.

97 MATHEMATICS AND COMPUTING

Popnet : computer vision based deep learning model for forecasting gridded population

Here, this study introduces Popnet, a deep learning model for forecasting 1 km-gridded populations, integrating U-Net, ConvLSTM, a Spatial Autocorrelation module and deep ensemble methods. Using spatial variables and population data from 2000 to 2020, Popnet predicts South Korea’s population trends by age groups (under 14, 15-64 and over 65) up to 2040. In validation, it outperforms traditional machine learning and state-of-the-art computer vision models. The output of this model discovered significant polarisation: population growth in urban areas, especially the capital region, and severe depopulation in rural areas. Popnet is a robust tool for offering significant insights to policymakers and related stakeholders about the detailed future population, which allows them to establish detailed, localised planning and resource allocations.

computer vision

SEASAT economic assessment. Volume 8: Ocean fishing case study

The potential application of SEASAT data with regard to ocean fisheries is discussed. Tracking fish populations, indirect assistance in forecasting expected populations and assistance to fishing fleets in avoiding costs incurred due to adverse weather through improved ocean conditions forecasts were investigated. Case studies on fisheries in the United States and Canada are cited.

Source record

Evaluation of the 2022 West Nile virus forecasting challenge, USA

Abstract Background West Nile virus (WNV) is the most common cause of mosquito-borne disease in the continental USA, with an average of ~1200 severe, neuroinvasive cases reported annually from 2005 to 2021 (range 386–2873). Despite this burden, efforts to forecast WNV disease to inform public health measures to reduce disease incidence have had limited success. Here, we analyze forecasts submitted to the 2022 WNV Forecasting Challenge, a follow-up to the 2020 WNV Forecasting Challenge. Methods Forecasting teams submitted probabilistic forecasts of annual West Nile virus neuroinvasive disease (WNND) cases for each county in the continental USA for the 2022 WNV season. We assessed the skill of team-specific forecasts, baseline forecasts, and an ensemble created from team-specific forecasts. We then characterized the impact of model characteristics and county-specific contextual factors (e.g., population) on forecast skill. Results Ensemble forecasts for 2022 anticipated a season at or below median long-term WNND incidence for nearly all (> 99%) counties. More counties reported higher case numbers than anticipated by the ensemble forecast median, but national caseload (826) was well below the 10-year median (1386). Forecast skill was highest for the ensemble forecast, though the historical negative binomial baseline model and several team-submitted forecasts had similar forecast skill. Forecasts utilizing regression-based frameworks tended to have more skill than those that did not and models using climate, mosquito surveillance, demographic, or avian data had less skill than those that did not, potentially due to overfitting. County-contextual analysis showed strong relationships with the number of years that WNND had been reported and permutation entropy (historical variability). Evaluations based on weighted interval score and logarithmic scoring metrics produced similar results. Conclusions The relative success of the ensemble forecast, the best forecast for 2022, suggests potential gains in community ability to forecast WNV, an improvement from the 2020 Challenge. Similar to the previous challenge, however, our results indicate that skill was still limited with general underprediction despite a relative low incidence year. Potential opportunities for improvement include refining mechanistic approaches, integrating additional data sources, and considering different approaches for areas with and without previous cases. Graphical Abstract

54 ENVIRONMENTAL SCIENCES

CovTransformer: A transformer model for SARS-CoV-2 lineage frequency forecasting

With hundreds of SARS-CoV-2 lineages circulating in the global population, there is an ongoing need for predicting and forecasting lineage frequencies and thus identifying rapidly expanding lineages. Accurate prediction would allow for more focused experimental efforts to understand pathogenicity of future dominating lineages and characterize the extent of their immune escape. Here, we first show that the inherent noise and biases in lineage frequency data make a commonly-used regression-based approach unreliable. To address this weakness, we constructed a machine learning model for SARS-CoV-2 lineage frequency forecasting, called CovTransformer, based on the transformer architecture. We designed our model to navigate challenges such as a limited amount of data with high levels of noise and bias. We first trained and tested the model using data from the UK and the USA, and then tested the generalization ability of the model to many other countries and US states. Remarkably, the trained model makes accurate predictions two months into the future with high levels of accuracy both globally (in 31 countries with high levels of sequencing effort) and at the US-state level. Our model performed substantially better than a widely used forecasting tool, the multinomial regression model implemented in Nextstrain, demonstrating its utility in SARS-CoV-2 monitoring. Assuming a newly emerged lineage is identified and assigned, our test using retrospective data shows that our model is able to identify the dominating lineages 7 weeks in advance on average before they became dominant. Overall, our work demonstrates that transformer models represent a promising approach for SARS-CoV-2 forecasting and pandemic monitoring.

60 APPLIED LIFE SCIENCES

Propagating synthetic populations with dynamic Bayesian networks: a framework for long-horizon demographic forecasting

This study presents a dynamic demographic microsimulator using dynamic Bayesian networks to forecast long–term changes in household and individual life events. Leveraging longitudinal Panel Study of Income Dynamics (PSID) data, two networks for individuals and households were modeled to simulate transitions in employment, income, education, marriage, childbirth, leaving the parental home, home ownership, mortality, and household formation or dissolution. Across 1,000 simulation runs spanning 24 years, household–level outcomes remain highly accurate and individual–level predictions reasonable. Although accuracy naturally declines with projection horizon, performance remains promising at both levels. This study addresses a key limitation of existing population synthesis models, which typically generate only a single static snapshot of the population. In conclusion, by introducing a framework that propagates cross-sectional outputs into the future, the microsimulator enables the tracking of demographic evolution over time, enhances realism in population-based simulations, and supplies credible inputs to agent-based travel demand models.

Demographic modeling

An integrated integral projection model ( IPM 2 ) to disentangle size‐structured harvest and natural mortality

Abstract Body size is one of the most important traits governing individual‐level demographic rates and modulating population‐level processes. Multiple size‐dependent demographic rates can simultaneously change population structure, so distinguishing their individual contributions to overall population dynamics remains a challenge. Disentangling size‐dependent harvest rates from other demographic rates is critical for assessing the impact of removal on populations of invasive species. Inference about invasive populations can be difficult, however, as observations are often collected opportunistically as part of removal programs, rather than experimentally designed. Yet accurate inference is essential for understanding the feasibility of population suppression and optimising management decisions. We develop an integrated integral projection model (IPM 2 ) that leverages the strengths of the integrated population model and integral projection model to enable inference about complex, size‐structured demographic rates from imperfect observations. We apply the IPM 2 in the context of invasive European green crab ( Carcinus maenas ), a species for which individual body size strongly regulates both the observation‐generating process and latent, population dynamics. The IPM 2 facilitates the distinct estimation of green crab size‐structured harvest and natural mortality rates, parameters for which no explicit data is collected and that are unidentifiable in component datasets of the integrated population model. The model represents how the green crab population changes over time, providing the first estimates of size‐structured abundance of this high‐priority species. By forecasting the stable size distribution and equilibrium population size under varying removal efforts, we demonstrate that extremely high levels of removal effort can reduce the equilibrium green crab population size. Yet these high mortality rates also shift the stable size distribution and increase the equilibrium abundance of smaller crabs, since size‐selective removal alters intraspecific interactions. The ecological outcome of this shift in size structure will be variable, as green crab size modulates only some of its interactions with other species. These results highlight the value of the IPM 2 framework for inferring complex population dynamics with information needs that outpace information in individual observational datasets, providing a path forward for accurate assessment of conservation programs.

Keller, Abigail G. [Department of Environment Scie

Gravitational waves and galaxies cross-correlations: a forecast on GW biases for future detectors

ABSTRACT Gravitational waves (GWs) have rapidly become important cosmological probes since their first detection in 2015. As the number of detected events continues to rise, upcoming instruments like Einstein Telescope (ET) and Cosmic Explorer (CE) will observe millions of compact binary mergers. These detections, coupled with galaxy surveys by instruments such as the Dark Spectroscopic Energy Instrument (DESI), Euclid, and the Vera Rubin Observatory, will provide unique information on the large-scale structure of the universe by cross-correlating GWs with the distribution of galaxies hosting them. In this paper, we focus on how cross-correlations constrain the clustering bias of GWs emitted by the coalescence of binary black holes (BBHs). This parameter links BBHs to the underlying dark matter distribution, hence informing us how they populate galaxies. Using a multitracer approach, we forecast the precision of these measurements under different survey combinations. Our results indicate that current GW detectors will have limited precision, with measurement errors as high as $\displaystyle \sim 50~{{\ \rm per\ cent}}$. However, third-generation detectors like ET, when cross-correlated with Legacy Survey of Space and Time (LSST) data, can improve clustering bias measurements to within 2.5 per cent. Furthermore, we demonstrate that these cross-correlations can enable a per cent-level measurement of the magnification lensing effect on GWs. Despite this, there is a degeneracy between magnification and evolution biases, which hinders the precision of both. This degeneracy is most effectively addressed by assuming knowledge of one bias or targeting an optimal redshift range of $\displaystyle 1 \lt z \lt 2.5$. Our analysis opens new avenues for studying the distribution of BBHs and testing the nature of gravity through large-scale structure.

Zazzera, Stefano (ORCID:0000000158979221)

EV Load Forecasting Guide: A Report by the Energy Systems Integration Group’s EV Load Forecasting Task Force

Forecasting electricity usage is a foundational planning activity for utilities, underpinning billions of dollars in grid investments that ensure system reliability. Historically, forecasting relied on trends in economic and population growth; however, transportation electrification presents a new and complex planning challenge. Unlike conventional loads, electric vehicle (EV) charging has relatively limited usage history. In addition, charging is driven by complex human behaviors, is mobile, and at the same time can concentrate geographically in ways that, without proper planning, can quickly overwhelm local distribution systems.

Giraldez, Julieta [Electric Power Engineers, Austi

Uncovering heterogeneous intercommunity disease transmission from neutral allele frequency time series

The COVID-19 pandemic has underscored the need for accurate epidemic forecasting to predict pathogen spread, evolution, and evaluate intervention strategies. Forecast reliability hinges on detailed knowledge of disease transmission across population segments, which may be inferred from contact surveys or mobility data. However, these indirect approaches make it difficult to estimate rare transmissions between socially or geographically distant communities. We show that the steep ramp-up of genome sequencing surveillance during the pandemic can be leveraged to directly identify transmission patterns between geographically defined communities. Our approach uses a hidden Markov model to infer the fraction of infections a community imports from others based on how rapidly allele frequencies in the focal community converge to those in the donor communities. Applying this method to SARS-CoV-2 sequencing data from England and the United States, we uncover networks of intercommunity transmission that reflect geographical relationships while exposing significant long-range interactions. The scaling of importation rate with distance is consistent across both countries, yet weaker than expected based on mobility data, highlighting limitations of indirect inference. We show that transmission patterns can change between waves of variants of concern and analyze how the inferred heterogeneity in intercommunity transmission impacts evolutionary forecasts. While applied here to geographically defined communities, our approach could be applied to those defined by other traits (e.g., age, socioeconomic status), provided time-series data can be stratified accordingly. Overall, our study highlights population genomic time series data as a crucial record of epidemiological interactions, which can be deciphered using tree-free inference methods.

Okada, Takashi [Department of Physics; University

Semi-Analytical Hierarchical Bayesian Inference of Nonlinear Model Structure in Stochastic Dynamics: Applied to Compartmental Models of Infectious Diseases

A Bayesian computational framework for parsimonious inference in stochastic nonlinear dynamical systems is presented. This framework enables the concurrent estimation of system states, time-varying parameters, time-invariant parameters, and the optimal sparsity structure of the model parameters. Because differential equation-based models are often simplified mechanistic or phenomenological representations, robust inference from noisy measurement data requires explicit treatment of model error and uncertainty. Model error and time-varying parameters can be represented as random processes, enabling inference while making minimal assumptions about the underlying sources of discrepancy and variability. Adopting stochastic differential equation representations affords the model significant flexibility, but can also render it susceptible to overfitting during statistical inversion, where the inferred model may track noise rather than the underlying signal. To alleviate the effects of overfitting and to enable the discovery of the optimal sparse representation of the time-invariant parameters, a Bayesian sparse learning algorithm is embedded within the framework. This sparse learning framework adopts an approximate hierarchical Bayesian setting defined by a series of semi-analytical expressions. The model structure inference framework is validated using a stochastic compartmental model for tracking and forecasting active cases of an infectious disease. Compartmental models describe population-level infectious disease dynamics through interactions among population fractions grouped by disease state. Mathematically, such models consist of a system of coupled ordinary differential equations. This example adopts an expressive compartmental model that includes multiple possible interactions between disease states, motivated by early uncertainty surrounding COVID-19 reinfection dynamics and their implications for long-term epidemic forecasting. The sparse learning exercise permits the inference of a priori unknown epidemiological dynamics from simulated public health data, discovering the nested compartmental model that optimizes the trade-off between average data-fit and model complexity. It is shown that inducing sparsity among the model parameters eliminates redundant interactions between compartments, equivalently revealing the optimal coupling structure between differential equations.

97 MATHEMATICS AND COMPUTING

The noise impact of proposed runway alternatives at Craig Airport

Four proposed runway expansion alternatives at Craig Airport in Jacksonville, Florida have been assessed with respect to their forecasted noise impact in the year 2005. The assessment accounts for population distributions around the airport and human subjective response to noise, as well as the distribution of noise levels in the surrounding community (footprints). The impact analysis was performed using the Airport-noise Levels and Annoyance Model (ALAMO), an airport community response model recently developd at Langley Research Center.

Deloach, R.

The large area crop inventory experiment: A major demonstration of space remote sensing

Strategies are presented in agricultural technology to increase the resistance of crops to a wider range of meteorological conditions in order to reduce year-to-year variations in crop production. Uncertainties in agricultral production, together with the consumer demands of an increasing world population, have greatly intensified the need for early and accurate annual global crop production forecasts. These forecasts must predict fluctuation with an accuracy, timeliness and known reliability sufficient to permit necessary social and economic adjustments, with as much advance warning as possible.

Macdonald, R. B.

Multiscale drivers of extreme southern California flooding: ENSO, MJO, North Pacific jet, and atmospheric rivers

Extreme rainfall and flooding, driven by a powerful atmospheric river (AR) and a persistent Madden-Julian Oscillation (MJO), hit Southern California in February 2024 during the 2023–2024 El Niño, affecting over 10 million people. ARs are key contributors to extreme rainfall and flooding along the U.S. West Coast. Although the AR-MJO link has been documented, its spatio-temporal variability remains a major forecasting and risk-management challenge. Combining precipitation, stream gauge and demographic data, we quantify the physical drivers and population exposure to this extreme event. Leveraging a Lagrangian MJO precipitation tracking algorithm, we unravel the multiscale interactions responsible for the AR’s development. El Niño favored a large, long-lived MJO that interacted with the North Pacific Jet (NPJ) over more than three weeks. The MJO convective outflow modulated the NPJ by inducing negative potential vorticity advection along the tropopause. The ensuing NPJ extension and acceleration induced explosive cyclogenesis, whose AR-driven moisture transport resulted in extreme rainfall.

Atmospheric dynamics

Studies in short haul air transportation in the California corridor: Effects of design runway length; community acceptance; impact of return on investment and fuel cost increases. Volume 2: Appendices

The development of a forecast model for short haul air transportation systems in the California Corridor is discussed. The factors which determine the level of air traffic demand are identified. A forecast equation for use in airport utilization analysis is developed. A mathematical model is submitted to show the relationship between population, employment, and income for indicating future air transportation utilization. Diagrams and tables of data are included to support the conclusions reached regarding air transportation economic factors.

Shevell, R. S.

Air Traffic Forecasting at the Port Authority of New York and New Jersey

Procedures for conducting air traffic forecasts with specific application to the Port Authority of New York and New Jersey are discussed. The procedure relates air travel growth to detailed socio-economic and demographic characteristics of the U.S. population rather than to aggregate economic data such as Gross National Product, personal income, and industrial production. Charts are presented to show the relationship between various selected characteristics and the use of air transportation facilities.

Augustine, J. G.

Relationship of physiography and snow area to stream discharge

The author has identified the following significant results. A comparison of snowmelt runoff models shows that the accuracy of the Tangborn model and regression models is greater if the test data falls within the range of calibration than if the test data lies outside the range of calibration data. The regression models are significantly more accurate for forecasts of 60 days or more than for shorter prediction periods. The Tangborn model is more accurate for forecasts of 90 days or more than for shorter prediction periods. The Martinec model is more accurate for forecasts of one or two days than for periods of 3,5,10, or 15 days. Accuracy of the long-term models seems to be independent of forecast data. The sufficiency of the calibration data base is a function not only of the number of years of record but also of the accuracy with which the calibration years represent the total population of data years. Twelve years appears to be a sufficient length of record for each of the models considered, as long as the twelve years are representative of the population.

Mccuen, R. H.