Statistical Failure Analysis of 6000 Residential Rooftop-Harvested Photovoltaic Connectors
Explore the source record for details and available documents.
SEARCH · Engineering Papers
Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.
Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.
Explore the source record for details and available documents.
Overview Code and data repository for NIL Manuscript. Documentation includes sequence processing examples and data analysis. Supplemental sequence processing and R statistical analysis for publication, which compares the microbiome of teosinte-B73 Near Isogenic Lines. Sample Data Amplicon sequence data for 16S rRNA genes, the fungal ITS2 region, and nitrogen-cycling functional genes are available through the NCBI Sequence Read Archive (SRA) under accession number PRJNA1042643(https://www.ncbi.nlm.nih.gov/bioproject/PRJNA1042643). Raw metabolomic data are available on Metabolomics Workbench, Project ID: PR002654. This study is available at the NIH Common Fund's National Metabolomics Data Repository (NMDR) website, the Metabolomics Workbench, https://www.metabolomicsworkbench.org where it has been assigned Study ID ST004211. The data can be accessed directly via its Project DOI: http://dx.doi.org/10.21228/M8KV8T.
Unambiguous identification of active sites in heterogeneous catalysis remains a major challenge, particularly for materials with ultrathin, chemically mixed surface layers. Here, we demonstrate a generalizable approach that combines time-of-flight secondary ion mass spectrometry (ToF-SIMS) with multivariate statistical analysis (principal component analysis [PCA] and multivariate curve resolution [MCR]) to resolve catalytically relevant motifs at the nanoscale. Using Ni electrodes as a model system, PCA distinguished hydroxide-enriched domains from oxide- and metal-rich regions, while MCR decomposed depth profiles and 3D images into hydroxide, oxide, and metallic layers with nanometer resolution. A unique secondary-ion fragment, NiO 3 H 3 − (m/z 108.94), emerged as a marker of hydroxide-rich environments and correlated with hydrogen evolution reaction (HER) activity across a series of Ni electrodes. Complementary density functional theory (DFT) calculations revealed that Ni(OH) 2 clusters adjacent to metallic Ni offer the most favorable water dissociation energetics, establishing the structural origin of the marker. While illustrated here for Ni-based HER, this workflow provides a broadly applicable framework to isolate and rank near-surface patterns that govern catalytic activity, thereby extending ToF-SIMS from a qualitative probe to a predictive tool for active site identification.
Although many studies have validated wave energy converter (WEC) numerical models against scaled prototype experimental data, there remains a notable lack of validation using data from full-scale deployed WECs. This paper compares two numerical models of Monterey Bay Aquarium Research Institute’s Wave Energy Converter (MBARI-WEC), a two-body point absorber with an electro-hydraulic power take-off system (PTO). The models are implemented in WEC-Sim/Simscape and Gazebo Simulator. A statistical analysis of the models was performed, and field results were obtained to compare the models’ accuracy in predicting the RMS piston velocity, RMS motor speed, and mean electric power compared to field data for 56 observations across varying sea states. The Gazebo model demonstrated a closer agreement across all three parameters for a majority of the observations. When compared to the field data, the Gazebo and WEC-Sim models exhibited average mean electric power overestimations of 13% and 22%, respectively.
Precise size and shape control in nanocrystal synthesis is essential for utilizing nanocrystals in various industrial applications, such as catalysis, sensing, and energy conversion. However, traditional ensemble measurements often overlook the subtle size and shape distributions of individual nanocrystals, hindering the establishment of robust structure–property relationships. In this study, we uncover intricate shape evolutions and growth mechanisms in Co 3 O 4 nanocrystal synthesis at a subnanometer scale, enabled by deep-learning-assisted statistical characterization. By first controlling synthetic parameters such as cobalt precursor concentration and water amount then using high resolution electron microscopy imaging to identify the geometric features of individual nanocrystals, this study provides insights into the interplay between synthesis conditions and the sizedependent shape evolution in colloidal nanocrystals. Utilizing population-wide imaging data encompassing over 441,067 nanocrystals, we analyze their characteristics and elucidate previously unobserved size-resolved shape evolution. This high-throughput statistical analysis is essential for representing the entire population accurately and enables the study of the size dependency of growth regimes in shaping nanocrystals. Our findings provide experimental quantification of the growth regime transition based on the size of the crystals, specifically (i) for faceting and (ii) from thermodynamic to kinetic, as evidenced by transitions from convex to concave polyhedral crystals. Additionally, we introduce the concept of an “onset radius,” which describes the critical size thresholds at which these transitions occur. This discovery has implications beyond achieving nanocrystals with desired morphology; it enables finely tuned correlation between geometry and material properties, advancing the field of colloidal nanocrystal synthesis and its applications.
En-route charging infrastructure for electric vehicles is critical to support transportation needs. These charging stations are likely to have high loads and especially sharp peak loads given fast charging capabilities needed to meet transportation schedules. In order to reduce both strain on distribution grid infrastructure and charging station operational costs, many stations are likely to employ behind the meter storage. This paper demonstrates a behind the meter storage sizing optimization that employs an open-source agent-based vehicle behavior model (BEAM) to determine the best sizing across many scenarios. This optimization and analysis is novel in that it examines how storage size impacts not only charging station cost and peak load, but also vehicle queue times. The optimization is also applied across a wide analysis region with sufficient diversity and numbers to provide novel statistical analysis of optimal sizes.
Proteoforms arising from posttranslational modifications, genetic polymorphisms, and RNA splice variants, play a pivotal role as the key drivers in biology. Thus, a comprehensive understanding of proteoforms is essential for unraveling the intricacies of biological systems and bridging the gap between genotype and phenotype. By analyzing whole proteins without digestion, top-down proteomics (TDP) provides a holistic view of the proteome and presents a next-generation approach for deciphering protein function, uncovering disease mechanisms, and advancing precision medicine. This Primer embarks on a journey into the world of TDP by encapsulating its historical context, underlying principles, recent advances, and an outlook on the future of TDP. The experimental section navigates instrumentation, sample preparation, intact protein separation, tandem mass spectrometry techniques, and data collection. Results decipher raw data, visualize intact protein spectra, unravel data analysis, and explain proteoform identification, characterization, and quantitation, as well as statistical analysis. Various applications of TDP spanning the human proteoform project, biomedical, biopharmaceutical, and clinical applications are described. These are complemented by discussions on measurement reproducibility, limitations, and a forward-looking perspective outlining uncharted waters where the field can advance, and potential exciting future applications of TDP.
Cislunar space, encompassing the region from geosynchronous orbit to beyond the Moon, is poised to become a cornerstone for future exploration, scientific discovery, and national security. Missions in this region, spanning durations from weeks to decades, require robust infrastructure and reliable transit capabilities. The complex gravitational influences of the Moon, Sun, and planets, along with thermal radiation from Earth and the Sun, lead to significant trajectory deviations, resulting in kilometer-scale errors within days. Leveraging the high-performance computing resources at Lawrence Livermore National Laboratory (LLNL), we have simulated one million high-fidelity cislunar trajectories, now publicly available via LLNL’s Green Data Oasis and the Unified Data Library. Generated using the open-source Space Situational Awareness Python package, these trajectories match the precision of commercial tools such as AGI’s Systems Tool Kit and NASA’s General Mission Analysis Tool. This data set is a valuable resource for reference, statistical analysis of cislunar orbit populations, and training machine learning models for rapid orbit classification with minimal observational input. Preliminary analysis reveals stable bands in Keplerian element space, particularly around five geosynchronous radii across a range of inclinations and eccentricities. Beyond this threshold, the Moon’s influence disrupts most unassisted orbits, though co-orbiting L4/L5 Lunar Trojans persist throughout the six-year simulation.
Wastewater treatment plants (WWTPs) are typically energy intensive, mainly due to the secondary treatment processes such as activated sludge (AS) for treatment of organics as well as nutrients like nitrogen. Nitrogen removal presents a big problem for WWTPs. The main form of nitrogen in wastewater is ammonium, and an AS process uses oxygen to convert ammonium into nitrite and nitrate which is then converted to nitrogen through denitrification process. During anaerobic digestion (AD), organic nitrogen gets degraded, resulting in an effluent stream (centrate) with a high nitrogen content, mostly in the form of ammonium. This contributes 15-30% of total nitrogen to the wastewater influent which further increases energy consumption for aeration. The project aims to transform this conventional municipal WWTPs into energy-neutral, resource-recovering facilities by integrating three core technologies: • Cloth Media Filtration (CMF) to replace conventional primary sedimentation (CPS) and increase the diversion of organics from the energy intensive secondary treatment to AD. This results in reduced energy demand for aeration in the secondary process while simultaneously increasing the biogas production in the anaerobic digesters. • Anerobic Digester to increase biogas and ammonia production. • Membrane Evaporation (ME) to recover ammonia from AD centrate and produce marketable fertilizer. The benefits of proposed WWTP process modifications were evaluated using techno economic analysis (TEA) and life cycle assessment (LCA). For CMF portion of the research a statistical analysis was employed to develop data-driven tools that could be used to enhance and optimize its performance in terms of energy savings and effluent quality. The main objective of this project is to reduce the energy demand for secondary treatment at municipal WWTPs by at least 50%, increase anaerobic digester (AD) biogas and ammonia production by 100% and 120%, respectively, and recover 90% of ammonia from the AD. Integrated CMF, AD, and ME was shown to work synergistically toward achieving these decarbonization targets through energy-positive treatment and fertilizer recovery techniques.
The past few years have seen the emergence of a wide array of novel techniques for analyzing high-precision data from upcoming galaxy surveys, which aim to extend the statistical analysis of galaxy clustering data beyond the linear regime and the canonical two-point (2pt) statistics. We test and benchmark some of these new techniques in a community data challenge named “Beyond-2pt,” initiated during the Aspen 2022 Summer Program “Large-Scale Structure Cosmology beyond 2-Point Statistics,” whose first round of results we present here. The challenge data set consists of high-precision mock galaxy catalogs for clustering in real space, in redshift space, and on a light cone. Participants in the challenge have developed end-to-end pipelines to analyze mock catalogs and extract unknown (“masked”) cosmological parameters of the underlying ΛCDM models with their methods. The methods represented are density-split clustering, nearest neighbor statistics, BACCO power spectrum emulator, void statistics, LEFTfield field-level inference using effective field theory (EFT), and joint power spectrum and bispectrum analyses using both EFT and simulation-based inference. In this work, we review the results of the challenge, focusing on problems solved, lessons learned, and future research needed to perfect the emerging beyond-2pt approaches. The unbiased parameter recovery demonstrated in this challenge by multiple statistics and the associated modeling and inference frameworks supports the credibility of cosmology constraints from these methods. The challenge data set is publicly available, and we welcome future submissions from methods that are not yet represented.
This dissertation explores factors influencing pooled rideshare (PR) adoption to provide actionable insights for transportation network companies (TNCs) and policymakers. PR allows travelers to share rides with unknown passengers, offering benefits such as cost reduction and congestion relief. However, adoption remains limited due to safety concerns, privacy issues, and trust in rideshare platforms. A national U.S. survey with 5,385 respondents examined transportation preferences and barriers to PR adoption. Exploratory and confirmatory factor analyses identified five key factors influencing PR consideration—safety, service experience, privacy, traffic/environment, and time/cost. Second factor analyses examined ways to optimize PR experiences, revealing four factors—comfort/ease of use, convenience, vehicle technology/accessibility, and passenger safety. Privacy concerns, for instance, using regression analysis, were found to reduce the likelihood of PR adoption by 77%, and convenience had the potential to increase it by 156%. The Pooled Rideshare Acceptance Model (PRAM), based on the Technology Acceptance Model, assessed the impact of these factors using the Structural Equation Model (SEM). Privacy, safety, trust, and convenience had a large effect (Cohen's f2 > 0.35) on PR acceptance, while multigroup analyses (PRAMMA) explored 16 demographic variables such as gender, generation, and income, emphasizing the need for tailored strategies. Based on all the statistical analysis and workshops using descriptive statistics, 95 actionable recommendations were made from the riders' perspective. Findings highlight the importance of customized services, user experience improvements, and policy interventions to enhance PR adoption. This dissertation provides a roadmap for future research and policy development, ensuring evidence-based, practical strategies to improve PR services in the U.S. and beyond.
This data package was generated to support the manuscript “Towards CONUS-Wide Machine Learning-Augmented Conceptually Interpretable Modeling of Catchment-Scale Precipitation-Storage-Runoff Dynamics.” It provides input files, model outputs, plotting data, scripts, notebooks, and documentation used to develop, evaluate, and reproduce Mass-Conserving Perceptron (MCP)-based hydrologic modeling experiments across 513 selected Catchment Attributes and Meteorology for Large-sample Studies in the United States (CAMELS-US) basins. The files are organized by modeling component and analysis purpose, including rainfall–runoff experiments, snow module experiments, coupled hydrologic-snow experiments, Long Short-Term Memory (LSTM) benchmark results, model skill metrics, initialization and epoch records, cell-state normalization files, Akaike Information Criterion (AIC)-based model comparison files, and data used to generate manuscript figures. Tabular files can be opened using standard spreadsheet software or Python/R data-analysis tools. Python scripts, Jupyter notebooks, and selected MATLAB scripts are included for model execution, postprocessing, plotting, and statistical analysis. Quality assurance and quality control were conducted through the source-data selection and modeling workflow. Meteorological forcing, streamflow, and static catchment attributes were derived from the CAMELS-US dataset, and snow water equivalent data were derived from the University of Arizona (UA) Snow Water Equivalent dataset. Selected basins and time periods were screened during the associated research workflow to avoid missing observations or poor-quality cases. Static geospatial features were processed primarily using Quantum Geographic Information System (QGIS) and Geospatial Data Abstraction Library (GDAL) workflows. Additional details are provided in the associated manuscript and documentation.
In this paper, we introduce PYSIMFRAC, an open-source python library for generating 3-D synthetic fracture realizations, integrating with fluid simulators, and performing analysis. PYSIMFRAC allows the user to specify one of three fracture generation techniques (Box, Gaussian, or Spectral) and perform statistical analysis including the autocorrelation, moments, and probability density functions of the fracture surfaces and aperture. This analysis and accessibility of a python library allows the user to create realistic fracture realizations and vary properties of interest. In addition, PYSIMFRAC includes integration examples to two different pore-scale simulators and the discrete fracture network simulator, dfnWorks. The capabilities developed in this work provides opportunity for quick and smooth adoption and implementation by the wider scientific community for accurate characterization of fluid transport in geologic media. We present PYSIMFRAC along with integration examples and discuss the ability to extend PYSIMFRAC from a single complex fracture to complex fracture networks.
ProteoMeter is a Python package that assists in the statistical analysis of global proteomics, protein post-translation modification (PTM), and limited proteolysis (LiP) data. It contains batch correction, normalization, and statistical testing methods, as well as functions that "roll up" peptide-level data to the single-site level. It has a robust user configuration system, allowing it to flexibly integrate different types of experiment designs. For basic usage, a simple configuration file provides the essential functionality. Advanced users have access to the entire statistical pipeline for fine-tuning analyses. Processed data is easily exported to many common spreadsheet and data-frame formats.
Benchmarking is essential for high-performance software development, particularly for monitoring performance across code iterations. This project focused on enhancing the benchmarking process for Lamellar, an asynchronous runtime for High-Performance Computing (HPC) systems developed at Pacific Northwest National Laboratory. Prior to this work, benchmark results were difficult to track and compare across code versions, presenting significant challenges in identifying performance regressions and long-term trends. The primary objective was to establish a systematic, reproducible approach for measuring performance and detecting regressions following code commits. Our methodology involved three key components: standardizing benchmark outputs, implementing data versioning, and developing analysis tools. We standardized the benchmark output format to JSON Line records containing specific fields (execution time, hardware specifications, and environmental variables). To address data management challenges, we evaluated several options and eventually chose a git repository dedicated to benchmark data. We developed a suite of Python tools that processed benchmark results, enriched them with metadata, and facilitated search in the repository. The resulting system enables more efficient filtering and comparison of performance metrics across commit histories, hardware configurations, and benchmark variants through a unified query interface. Our implementation reduces computational overhead by first checking for existing results through configuration matching before initiating new benchmark runs, thereby conserving resources. The system has been validated by Lamellar developers. It organizes results by benchmark type and build configurations for efficient retrieval. Future developments include a planned Large Language Model interface for predicting benchmark performance, incorporating the criterion package for statistical analysis, which will enable automated detection of statistically significant performance changes, and integration with continuous integration pipelines. Despite these enhancements being reserved for future work, this project has successfully provided the Lamellar development team with a framework for maintaining consistent performance standards and identifying optimization opportunities across workloads and hardware environments.
This dataset contains solid-state 13C NMR data and atomistic molecular dynamics simulation files supporting the study of nanoscale secondary cell wall architecture across 13 genetically diverse Populus trichocarpa genotypes grown under uniform greenhouse conditions in 13C-enriched CO2 atmospheres (~89% 13C enrichment).The dataset contains two collections of solid-state 13C NMR data. (1) 200 MHz data (Bruker Avance III HD, 4 mm HX probe, 10 kHz MAS): raw Bruker TopSpin experiment folders and DMFIT-exported ascii spectra for selective and non-selective 1D 13C-13C spin diffusion experiments (3000 ms mixing) used to quantify inter-polymer spatial proximities, and short-mixing (1 ms) reference spectra used for polymeric abundance quantification by spectral deconvolution. (2) 600 MHz data (Bruker Avance III, 1.6 mm PhoenixNMR HXY probe, 30 kHz MAS): raw Bruker TopSpin experiment folders containing 2D CORD, 2D CP-INADEQUATE, and 13C/1H relaxation (T1, T1rho) experiments for all 13 genotypes, with processed Excel workbooks per experiment type. Molecular dynamics simulation code, coordinate files, and analysis scripts (NAMD/CHARMM/Python) for six atomistic cell wall models are included. Summarized ssNMR data are compiled into a single excel file and subjected to statistical analysis. Multivariate analysis code (PCA, Pearson correlation) and summary data are provided as excel worksheets and Jupyter notebooks (Python 3).
Atom probe tomography (APT) has been utilized to investigate the microstructure of two model borosilicate glasses designed to understand the solubility limits of phosphorous pentoxide (P 2 O 5 ). This component is found in certain high-level radioactive defence wastes destined for vitrification, where phase separation can potentially lead to a number of issues relating to the processing of the glass and its long-term chemical and structural stability. The development of suitable focused ion beam (FIB)-preparation routes and APT analysis conditions were initially determined for the model glasses, before examining their detailed microstructures. In a 3.0 mol% P 2 O 5 -doped glass, both visual inspection and sensitive statistical analysis of the APT data show homogeneous microstructures, while raising the content to 4.0 mol% initiates the formation of phosphorus-enriched nanoscale precipitates. This study confirms the expected inhomogeneities and phase separation of these glasses and offers routes to characterizing these at near-atomic scale resolution using APT.
The dynamic behaviour of floating offshore wind turbines (FOWTs) involves complex interactions of multivariate loads from wind, waves, and currents, which result in complex motion characteristics. Although methods for analysing global motion responses are well-established, the time- and location-dependent kinematics remain underexplored. This paper investigates the instantaneous centre of rotation (ICR), a point of zero velocity at a time instance of general plane motion. Understanding and strategically positioning the ICR can reduce the dynamic motion in critical structural locations, enhancing the performance and structural robustness of FOWTs. The paper presents a method for computing the ICR using time-domain simulation results and proposes a statistical analysis approach suitable for design studies. Building on prior research, it examines the sensitivity of the ICR to external loading and design features, providing insights into how these factors influence motion response and how the motion response influences the statistics of the ICR, structural loads, and other performance metrics of interest. The study explores two FOWT configurations, a spar and a semisubmersible, identifying design variables that most effectively control the ICR statistics and identifying the ICR statistics most correlated with the responses of interest. Finally, through two case studies, we demonstrate how to apply these new insights in a practical design scenario. By adjusting the design variables most correlated with the ICR (fairlead vertical position and centre of mass for the spar and mooring line length and offset column diameter for the semisubmersible), we successfully modified the designs of the floating support structures to reduce the loads in the mooring lines, tower base, and blade roots, improving the ultimate strength and fatigue characteristics compared to the original designs.