Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “synthetic population”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Initial Mobility Analysis for ORNL VA-EDH Synthetic Populations

Travel burdens are a major barrier to healthcare access among US Veteran patient populations, particularly those residing in rural areas. Spatial accessibility to points of care for US Veteran populations is commonly assessed in two ways. The first approach uses open data from the US Census to represent collective travel burdens, for example the distance between population-weighted census tract centroids and VHA points of care. The second approach uses restricted-access VHA patient data to measure travel costs (e.g., distance, time) for accessing points of care with respect to geolocated patient addresses and real or approximated transportation networks. While the advantage of the open data approach lies in its reproducibility, it has notable limitations in its tendency to infer individual travel behavior from aggregate population characteristics, a problem known as ecological fallacy. Conversely, while the patient data approach is able to account for individual travel behavior, its ability to account for localized access disparities (e.g., a neighborhood with exceptionally high transportation costs) and patient demographics is limited as protecting individual patient data requires their storage in closed systems with limited capacity for adequately modeling real-world travel patterns or for supplementing patient attributes. Additionally, the patient data approach cannot account for veterans who are not enrolled in the VHA system but who may be eligible for care. These challenges limit the ability to perform “what if” analyses on the effects of place-specific interventions on veteran populations with high access barriers to healthcare. To address these challenges, we explore the application of realistic synthetic populations to examine travel burdens and spatial accessibility issues among veteran patient populations. Synthetic populations provide a virtual, individually-resolved and cross-sectional representation of the veteran patient population that enables investigation of spatial access to points of care in ways in which aggregate data and patient data do not. First, synthetic populations allow one to directly assess how individuals access points of care, from synthesized residential locations to outpatient facilities on real-world transportation networks. Modeling access to points of care at the individual scale addresses the ecological fallacy problem associated with using aggregated census data to represent veteran populations and patterns of movement. Second, synthetic populations provide a means of completely representing an area’s veteran population using only publicly available, anonymized census microdata from the American Community Survey (ACS) to ensure the privacy of real-world individuals. Generating synthetic populations from the ACS also expands descriptive characteristics beyond what patient data typically offers to include socio-demographic, economic, housing, and mobility attributes. More detailed profiles of both VHA patient populations and veterans not enrolled in the VA system will provide a comprehensive picture of groups that may benefit from interventions or outreach. As an initial exercise for using synthetic populations to measure veteran travel burdens to VA care, we apply Oak Ridge National Laboratory’s (ORNL) UrbanPop capability to generate a series of synthetic VHA patient populations for 9 Veterans Integrated Services Networks (VISN) market areas in 9 Census Divisions across the continental United States, which are listed in Table 1. We use UrbanPop to produce synthetic populations for the VISN markets selected for each US Census Division, then assign VA outpatient clinic destinations to synthetic VHA patients based on travel about each VISN market’s road network. To demonstrate using the synthetic populations to evaluate healthcare travel burdens, we compare the time-based impedance between simulated home locations and VA outpatient clinics in each VISN market. We then perform validation exercises on the synthetic populations with respect to neighborhood (block group) demographic composition as well as patient mobility, comparing aggregate origin-destination statistics for the synthetic population to outpatient visits available in restricted patient data from the VA’s Corporate Data Warehouse (CDW) database.

97 MATHEMATICS AND COMPUTING↗

Scalable Generation of High-fidelity Synthetic Population Ensembles

Used within social simulations, synthetic population ensembles enable uncertainty quantification (UQ) methods for obtaining more robust model inference and prediction. A synthetic population ensemble is a series of plausible virtual reconstructions of an area’s population at the granularity of people and residences, generated stochastically to preserve privacy of the source population survey’s respondents. In this paper, we demonstrate the production of large synthetic population ensembles for the U.S. via Oak Ridge National Laboratory’s UrbanPop framework to support modeling of high spatial resolution energy affordability metrics from nationwide social surveys in collaboration with the fusionACS project. The study involves two scenarios: creating ensembles for (1) 17 U.S. metropolitan areas in 2019 and (2) full U.S. Census Divisions in 2023, with each scenario consisting of 41 population instances (a base realization and 40 replicates). To accomplish this task at scale, we configured an integrated system within a research cloud, comprised of virtual containerizations, GPU-enhanced functionality, and orchestrated deployments of UrbanPop’s maturing Likeness Python ecosystem. Results demonstrate we maintained high-fidelity approximations of residential totals by areas of interest and the demographic characteristics of neighborhoods while reducing manual workflow burdens. Finally, we discuss plans to fine-tune and further develop our automated workflows for truly distributed job orchestration to increase computational efficiency, as well as provide an outlook for broadening applications of the ensembles.

Cluster computing↗

Producing High-fidelity Synthetic Population Ensembles at Scale

Used within social simulations, synthetic population ensembles enable uncertainty quantification (UQ) methods for obtaining more robust model inference and prediction. A synthetic population ensemble is a series of plausible virtual reconstructions of an area’s population at the granularity of people and residences, generated stochastically to preserve privacy of the source population survey’s respondents. In this paper, we demonstrate the production of large synthetic population ensembles for the US via Oak Ridge National Laboratory’s UrbanPop framework to support modeling of high spatial resolution energy affordability metrics from nationwide social surveys in collaboration with the fusionACS project. Our initial task involves creating ensembles for 17 US metropolitan areas, each consisting of 41 population instances (a base realization and 40 replicates). To accomplish this task at scale, we configured an integrated system comprised of a research cloud, virtual containerization, GPU-enhanced functionality, and a dual API/CLI to interact with UrbanPop’s maturing Likeness Python ecosystem. We observe a reduction in theoretical execution time while maintaining high-fidelity approximations of residential totals by metropolitan area and the demographic characteristics of neighborhoods. We discuss expansion of our approach to produce synthetic population ensembles for the entire US, particularly plans to establish automated workflows for job orchestration to increase computational efficiency, as well as provide outlook for broadening applications of the ensembles.

Gaboardi, James [ORNL] (ORCID:0000000247766826)↗

Downscaling Synthetic Populations to Realistic Residential Locations

High-fidelity pattern of life (PoL) models require realistic origin points for predictive trip modeling. This paper develops and demonstrates a reproducible method using open data to match synthetic populations generated from census surveys to plausible residential locations (building footprints) based on housing attributes. This approach presents promise over extant methods based on housing density, particularly in small neighborhood areas with heterogeneous land-use.

Tuccillo, Joe↗

Propagating synthetic populations with dynamic Bayesian networks: a framework for long-horizon demographic forecasting

This study presents a dynamic demographic microsimulator using dynamic Bayesian networks to forecast long–term changes in household and individual life events. Leveraging longitudinal Panel Study of Income Dynamics (PSID) data, two networks for individuals and households were modeled to simulate transitions in employment, income, education, marriage, childbirth, leaving the parental home, home ownership, mortality, and household formation or dissolution. Across 1,000 simulation runs spanning 24 years, household–level outcomes remain highly accurate and individual–level predictions reasonable. Although accuracy naturally declines with projection horizon, performance remains promising at both levels. This study addresses a key limitation of existing population synthesis models, which typically generate only a single static snapshot of the population. In conclusion, by introducing a framework that propagates cross-sectional outputs into the future, the microsimulator enables the tracking of demographic evolution over time, enhances realism in population-based simulations, and supplies credible inputs to agent-based travel demand models.

Demographic modeling↗

DEMOS (Demographic Microsimulator Tool for Longitudinal Synthetic Population) [SWR-25-135] related to NLR SWR-26-076

The Demographic Microsimulator (DEMOS) is an agent-based simulation framework used to model the evolution of population demographic characteristics and lifecycle events, such as education attainment, marital status, and other key transitions. DEMOS modules are designed to capture the interdependencies between short-term and long-term lifecycle events, which are often influential in downstream transportation and land-use modeling. A key feature of DEMOS is its ability to track changes in an agent’s demographic status from year t to year t + 1. This structure allows the model to evolve populations over any user-defined time horizon. As a result, DEMOS is well suited for analyzing medium- and long-term transportation-related decisions, including household vehicle transactions (e.g., purchasing, selling, or replacing vehicles) and work location choices. Core features of DEMOS include the modeling of more than ten lifecycle events, behaviorally realistic patterns informed by long-running panel data, explicit representation of interdependencies among lifecycle processes, and a flexible, modular simulation architecture. A technical memorandum describing DEMOS is available here. The memorandum provides an overview of the framework’s functionality, model structure, input and output data, and its applications in transportation planning and broader policy analysis contexts. Interested readers are also encouraged to consult the paper listed below for additional details on the DEMOS methodology. Sun, Bingrong, Shivam Sharda, Venu M. Garikapati, Mohamed Amine Bouzaghrane, Juan Caicedo, Srinath Ravulaparthy, Isabel Viegas de Lima, Ling Jin, C. Anna Spurlock, and Paul Waddell. "Demographic Microsimulator for Integrated Urban Systems: Adapting Panel Survey of Income Dynamics to Capture the Continuum of Life." Transportation Research Record (2025): 03611981251333339.

Sun, Bingrong [National Laboratory of the Rockies ↗

UrbanPop: A spatial microsimulation framework for exploring demographic influences on human dynamics

Ensuring the social equity of planning measures in social systems requires an understanding of human dynamics, particularly how individual relationships, activities, and interactions intersect with individual needs. Spatial microsimulation models (SMSMs) support planning for human security goals by representing human dynamics through realistic, georeferenced synthetic populations, that a) provide a complete representation of social systems while b) also protecting individual privacy. In this paper, we present UrbanPop, an open and reproducible SMSM framework for analysis of human dynamics with high spatial, temporal, and demographic resolution. UrbanPop creates synthetic populations of demographically detailed worker and student agents, positioning them first at probable nighttime locations (home), then moving them to probable daytime locations (work/school). Summary aggregations of these populations match the granular detail available at the census block group level in the American Community Survey Summary File (SF), providing realistic approximations of the actual population. UrbanPop users can select particular demographic traits important in their application, resulting in a highly tailored agent population. We first lay out UrbanPop's baseline methodology, including population synthesis, activity modeling, and diagnostics, then demonstrate these capabilities by developing case studies of shifting population distributions and high-risk populations in Knox County, TN during the global COVID-19 pandemic.

60 APPLIED LIFE SCIENCES↗

Agent-Based Model of Combined Community- and Jail-Based Take-Home Naloxone Distribution

Importance Opioid-related overdose accounts for almost 80 000 deaths annually across the US. People who use drugs leaving jails are at particularly high risk for opioid-related overdose and may benefit from take-home naloxone (THN) distribution. Objective To estimate the population impact of THN distribution at jail release to reverse opioid-related overdose among people with opioid use disorders. Design, Setting, and Participants This study developed the agent-based Justice-Community Circulation Model (JCCM) to model a synthetic population of individuals with and without a history of opioid use. Epidemiological data from 2014 to 2020 for Cook County, Illinois, were used to identify parameters pertinent to the synthetic population. Twenty-seven experimental scenarios were examined to capture diverse strategies of THN distribution and use. Sensitivity analysis was performed to identify critical mediating and moderating variables associated with population impact and a proxy metric for cost-effectiveness (ie, the direct costs of THN kits distributed per death averted). Data were analyzed between February 2022 and March 2024. Intervention Modeled interventions included 3 THN distribution channels: community facilities and practitioners; jail, at release; and social network or peers of persons released from jail. Main Outcomes and Measures The primary outcome was the percentage of opioid-related overdose deaths averted with THN in the modeled population relative to a baseline scenario with no intervention. Results Take-home naloxone distribution at jail release had the highest median (IQR) percentage of averted deaths at 11.70% (6.57%-15.75%). The probability of bystander presence at an opioid overdose showed the greatest proportional contribution (27.15%) to the variance in deaths averted in persons released from jail. The estimated costs of distributed THN kits were less than $\$$15 000 per averted death in all 27 scenarios. Conclusions and Relevance This study found that THN distribution at jail release is an economical and feasible approach to substantially reducing opioid-related overdose mortality. Training and preparation of proficient and willing bystanders are central factors in reaching the full potential of this intervention.

Tatara, Eric [Argonne National Laboratory (ANL), A↗

An Interpretable Index of Social Vulnerability to Environmental Hazards

Index-based measures of social vulnerability to environmental hazards are commonly modeled from composites of population-level risk factors. These models overlook individual context in communities' experiences of environmental hazards, producing metrics that may hinder spatial decision support for mitigating and responding to hazards. This paper introduces an interpretable, high-resolution model for generating an individual-oriented social vulnerability index (IOSVI) for the United States built on synthetic populations that couples individual and social determinants of vulnerability. The IOSVI combines an individual vulnerability index (IVI) that ranks individuals in an area’s synthetic population based on intersecting risk factors, with a social vulnerability index (SVI) based on the population’s cumulative distribution of IVI scores. Interpretability of the IOSVI procedure is demonstrated through examples of national, metropolitan, and neighborhood (census tract) level spatial variation in index scores and IVI themes, as well as an exploratory analysis examining risk factors affecting a specific sub-population (military veterans) in areas of high social and environmental vulnerability.

Tuccillo, Joe↗

Demographic Microsimulator for Integrated Urban Systems: Adapting Panel Survey of Income Dynamics to Capture the Continuum of Life

Agent-based models (ABMs) in transportation modeling simulate activity and travel decisions at the disaggregate level of households and individuals. To do this, ABMs require detailed and realistic information on agents’ socioeconomic and demographic characteristics. Various synthetic population generators have been proposed to address this need. However, most of those currently in practice are cross-sectional in nature and do not account for the dynamics within households and individuals as they progress through life events over time. This is a major shortcoming, as literature has shown that transportation decisions are affected by the transition between and co-occurrence of life cycle events. While some demographic evolution simulators have been proposed to address this issue, they are developed using cross-sectional data and capture only a small set of life cycle events and their interdependence. Addressing these drawbacks, we propose a demographic microsimulator (DEMOS) that captures the “continuum of life” by considering a range of household- and individual-level life cycle events. DEMOS is developed using the Panel Survey of Income Dynamics, one of the world’s longest-running longitudinal surveys. The DEMOS submodels consider key life cycle events that are influenced by agents’ demographic variables. DEMOS is applied to evolve the population of the San Francisco Bay Area over a 9-year horizon. Results demonstrate how DEMOS generates life trajectories and how DEMOS outputs match the observed demographic trends. DEMOS is expected to enable longitudinal analysis in the context of ABMs and expand ABMs analyses relating to dynamic processes such as household-level vehicle transactions.

Demographic evolution↗

Global Properties of Local Star Forming Galaxies (ADP 2000)

We performed an archival study of the Hopkins Ultraviolet Telescope (HUT) Astro-2 database. Nineteen spectra of star-forming regions and starburst galaxies were retrieved, reprocessed, and analyzed. The spectra cover the wavelength region 912- 1800 A, providing access to the domain of peak luminosity from a young stellar population. We created an atlas of galaxy spectra documenting the continuum and line properties with an emphasis on the relatively unexplored spectral region below 1200 A. The dust obscuration law was derived from a comparison of the HUT spectra with synthetic population models. The law is similar to the commonly adopted starburst reddening curve at longer wavelengths and approaches the Milky Way law near the Lyman break. A simple power-law parameterization is given, which allows users to express the reddening law in terms of the stellar or nebular color excess at ultraviolet or optical wavelengths. We studied the effect of time-dependent dust obscuration on synthetic ultraviolet line profiles of a young stellar population. If the youngest and most massive stars are more obscured than the older, less massive stars, the C IV 1550 and other stellar wind lines are significantly diluted with respect to a simple foreground screen model for the dust. We propose to use stellar wind lines as a probe of the dust-obscuration model instead of the previously employed nebular emission lines. Since purely stellar diagnostics are utilized, uncertain assumptions on the nebular properties are unnecessary. Photoionization models demonstrate that the C IV 1550 emission is typically dominated by stellar winds and nebular contamination is negligible. A first comparison with the galaxy sample observed with the Hopkins Ultraviolet Telescope favors a dust geometry affecting ionizing and nonionizing stars equally. We point out the need for higher quality data for a more rigorous comparison. The Hubble Space Telescope is capable of obtaining such data in the future.

Leitherer, Claus↗

UrbanPop 2019 Baseline Population

A synthetic residential baseline population for the United States at the census block group level based on the American Community Survey 2015-2019 5-Year Estimates and described by residential, demographic, social, economic, student, mobility, and housing characteristics.

Tuccillo, Joe [ORNL] (ORCID:0000000259300943)↗

Simulating nationwide coupled disease and fear spread in an agent-based model

Human cognitive responses, behavioral responses, and disease dynamics co-evolve over the course of any disease outbreak, and can result in complex feedbacks. We present a dynamic agent-based model that explicitly couples the spread of disease with the spread of fear surrounding the disease, implemented within the EpiCast simulation framework. EpiCast models transmission within a realistic synthetic population, capturing individual-level interactions. In our model, fear propagates through both in-person contact and broadcast media, prompting individuals to adopt protective behaviors that reduce disease spread. In order to better understand these coupled dynamics, we create and compare a range of compartmental models to ensure that introducing additional disease states does not prevent the emergence of multiple waves in these simpler models. Additionally, we compare a range of behavioral scenarios within EpiCast, varying the level and intensity of fear and behavior change. Our results show that the addition of asymptomatic, exposed, and pre-symptomatic disease states can impact both the rate at which an outbreak progresses and its overall trajectory in compartmental models. In EpiCast, the combination of non-local fear spread via broadcasters and strong behavioral responses by fearful individuals generally leads to multiple epidemic waves, an outcome that occurs only within a narrow parameter range when fear spreads purely through local contact. Accounting for the coupled spread of fear and disease is critical for understanding disease dynamics and designing timely, targeted responses to emerging infectious threats.

60 APPLIED LIFE SCIENCES↗

Platform for Integrated Land use And Transportation Experiments and Simulation (PILATES) v1.0

PILATES allows for flexibly and at-scale coupling of multiple models to allow for multi-scale and multi-resolution simulation of regional-scale transport networks. In particular, it couples the MATSim-derived transportation modeling framework for Behavior, Energy, Autonomy and Mobility (BEAM) with other models operating at different time scales. Rather than tightly coupling supply and demand models using shared agents and memory within the same software process, PILATES orchestrates different model runs in a containerized framework. This structure requires passing information from the demand models to BEAM in the format of a synthetic population and agent plans, and from BEAM to the demand models in terms or origin/destination tables (also known as "skims"). This allows it to take advantage of the behavioral sophistication of existing activity-based models as well as the reinforcement learning structure of MATSim replanning and adopted by BEAM, in a way that requires minimal changes to existing models. It also takes advantage of the computational performance of BEAM, which allows for simulations with millions of agents to complete in reasonable time as well as allowing for detailed mechanistic simulation of the operation of on-demand modes.

Needell, Zachary↗

Firm Synthesizer and Supply-chain Simulator (SynthFirm) v1.0

SynthFirm is a large-scale agent-based freight demand model which generates a complete synthetic population of firms in the U.S. and the business-to-business commodity flows between them. Using publicly available data sources as inputs, SynthFirm simulates detailed firm and fleet characteristics, commodity production and consumption, formation of supply chains, and selection of shipping modes, all of which are essential drivers of commodity flow at a disaggregate level.

Xu, Xiaodan↗

Behavior, Energy, Autonomy, Mobility Modeling Framework (BEAM) v1.0

The Behavior, Energy, Autonomy, and Mobility (BEAM) model is an integrated, agent-based travel demand simulation framework. Individual agents express preferences through a utility- maximizing evolutionary algorithm that minimizes each individual’s cost and time spent traveling via diverse modal options, including the competition for scarce supply resources such as parking spaces and charging infrastructure. BEAM simulates the essential elements that compose a dynamic transportation system. From the road network, parking and charging infrastructure, to the transit system and a synthetic population with plans and preferences, the virtual system is an amalgamation of multiple spatially resolved layers that together represent an integrated transportation system. BEAM is an extension to the MATSim (Multi-Agent Transportation Simulation) model, where agents employ reinforcement learning across successive simulated days to maximize their personal utility through plan mutation (exploration) and selecting between previously executed plans (exploitation). The BEAM model shifts some of the behavioral emphasis in MATSim from across-day planning to within- day planning, where agents dynamically respond to the state of the system during the mobility simulation. In BEAM, agents can plan across all major modes of travel including driving, walking, biking, transit, and demand-responsive ride hailing. It is designed to integrate with other open source transportation models, such as ActivitySim.

Lazarus, Jessica↗

Firm Synthesizer and Supply-chain Simulator (SynthFirm) v2.0

SynthFirm is a national-scale agent-based freight demand model which generates a complete synthetic population of firms in the U.S. and the business-to-business commodity flows between them. Using publicly available data sources as inputs, SynthFirm simulates detailed firm and fleet characteristics, commodity production and consumption, formation of supply chains, and selection of shipping modes, all of which are essential drivers of commodity flow at a disaggregate level. The SynthFirm 2.0 version includes national commercial vehicle fleet generation, international trade simulation and automized model validation pipeline, which allows seemless deployment across the nation and build a comprehensive freight inventories at national scale or for selected region.

Yang, Hung-Chia [Lawrence Berkeley National Labora↗

Evaluating the Impacts of Autonomous Electric Vehicles Adoption on Vehicle Miles Traveled and CO2 Emissions

Autonomous electric vehicles (AEVs) can potentially revolutionize the transportation landscape, offering a safer, contact-free, easily accessible, and more eco-friendly mode of travel. Prior to the market uptake of AEVs, it is critical to understand the consumer segments that are most likely to adopt these vehicles. Beyond market adoption, it is also important to quantify the impact of AEVs on broader transportation systems and the environment, such as impacts on the annual vehicle miles traveled (VMT) and greenhouse gas (GHG) emissions. In this pilot study, using survey data, a statistical model correlating AEV adoption intention and socioeconomic and built environment attributes was estimated, and a sensitivity analysis was conducted to understand the importance of factors impacting AEV adoption. We found that the market segments range from early adopters who are wealthy, technologically savvy, and relatively young to non-adopters who are more cautious to new technologies. This is followed by a synthetic population microsimulation of market penetration for the San Francisco Bay Area. With five household vehicle replacement scenarios, we assessed the annual VMT and tailpipe carbon dioxide (CO2) emissions change associated with vehicle replacement. It is found that adopting AEVs can potentially reduce more than 5 megatons of CO2 yearly, which is approximately 30% of the total CO2 emitted by internal combustion engine (ICE) cars in the region.

33 ADVANCED PROPULSION SYSTEMS↗