Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “web - based tool”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

66 records · Page 4

bmdrc: Python package for quantifying phenotypes from chemical exposures with benchmark dose modeling

Though chemical exposures are known to potentially have negative impacts on health, including contributing to chronic diseases such as cancer, the quantitative contribution of risk is not fully understood for every chemical. A commonly used approach to quantify levels of risk is to measure the proportion of organisms (such as a total number of zebrafish on a plate or mice in a cage) with abnormal behavioral responses or morphology at increasing concentrations of chemical exposure. A particular challenge with processing the proportional data from these assays is the appropriate estimation of chemical concentration levels that result in malformations or acute toxicity, as these values typically vary between experimental measurements. The recommended approach by the Environmental Protection Agency (EPA) is to fit benchmark dose curves with specific filters and model fitting steps, which are crucial to properly processing the proportional data. Several tools exist for the fitting of benchmark dose response curves, but none are standalone Python libraries built to process both morphological and behavioral data as proportions with all the EPA recommended filters, filter parameters, models, and model parameters. Thus, here we present the benchmark dose response curve (bmdrc) Python library, which was built to closely follow these EPA guidelines with helpful visualizations of filters and fitted model curves, and reports for reproducibility purposes. bmdrc is open-source and has demonstrated utility as a support package to an existing web portal for information on chemicals (https://srp.pnnl.gov). Our package will support any toxicology analysis where the response is a proportional value at increasing levels of a concentration of a chemical or chemical mixture.

Superfund↗

GraphAide: Advanced Graph-Assisted Query and Reasoning System

Curating knowledge from multiple siloed sources that contain both structured and unstructured data is a major challenge in many real-world applications. Pattern matching and querying represent fundamental tasks in modern data analytics that leverage this curated knowledge. The development of such applications necessitates overcoming several research challenges, including data extraction, named entity recognition, data modeling, and designing query interfaces. Moreover, the explainability of these functionalities is critical for their broader adoption. The emergence of Large Language Models (LLMs) has accelerated the development lifecycle of new capabilities. Nonetheless, there is an ongoing need for domain-specific tools tailored to user activities. The creation of digital assistants has gained considerable traction in recent years, with LLMs offering a promising avenue to develop such assistants utilizing domain-specific knowledge and assumptions. In this context, we introduce an advanced query and reasoning system, GraphAide, which constructs a knowledge graph (KG) from diverse sources and allows to query and reason over the resulting KG. GraphAide harnesses both the KG and LLMs to rapidly develop domain-specific digital assistants. It integrates design patterns from retrieval augmented generation (RAG) and the semantic web to create an agentic LLM application. GraphAide underscores the potential for streamlined and efficient development of specialized digital assistants, thereby enhancing their applicability across various domains.

Purohit, Sumit [BATTELLE (PACIFIC NW LAB)] (ORCID:↗

rcsb-api : Python Toolkit for Streamlining Access to RCSB Protein Data Bank APIs

The Protein Data Bank (PDB) was founded in 1971 as the first open-access digital data resource in biology to serve as the single global archive for three-dimensional (3D) macromolecular structure data. Current PDB holdings exceed 230,000 experimentally determined structures of proteins, nucleic acids, viruses, and macromolecular machines. The RCSB Protein Data Bank RCSB.org research-focused web portal facilitates search, analyses, and visualization of every PDB structure along with more than one million Computed Structure Models from AlphaFold DB and the ModelArchive. It is powered by a set of publicly available Application Programming Interfaces (APIs) that both support RCSB.org users and provide programmatic access to PDB data. Given the breadth and levels of granularity encompassed in this rich data collection, efficiently accessing the information programmatically may be challenging for new users. RCSB PDB has developed a Python software package, rcsb-api , that facilitates easy and efficient use of RCSB PDB APIs within a Python environment. This software tool is designed to streamline access to the extensive corpus of data housed within the PDB, enabling researchers to search, retrieve, and analyze 3D biostructure data seamlessly. Its use will accelerate research in structural biology, molecular biology and biochemistry, drug discovery, and bioinformatics by providing more efficient tools for data integration and analysis. The new toolkit is available on GitHub (github.com/rcsb/py-rcsb-api) and published to the public Python package repository (PyPI) to foster wider usage and support basic and applied research in fundamental biology, biomedicine, and the energy sciences.

FAIR principles↗

Extraction and Analysis of Time Series Data from Building Automation Systems Using Large Language Models

Semantic schemas like Haystack 4, Brick and ASHRAE standard 223 enable the structured, standardized, and machine-readable representation of building data, facilitating interoperability, data integration, and advanced analytics. However, extracting information from these models requires specialized expertise in SPARQL and other programming languages, skills that are not commonly found among building professionals. Recent advancements in Large Language Models (LLMs), such as ChatGPT, enable the construction of queries using natural language, making it easier for individuals to interact with these systems in a manner that resembles everyday speech. However, these methods have not yet been tested on building semantic ontologies. This paper introduces a novel workflow and tool for enabling users to ask questions about a specific building's data, using natural language and receive answers automatically generated by GPT-4o. Our approach integrates semantic ontologies with advanced LLM capabilities to automate three critical steps: (1) generating SPARQL queries to retrieve time series references from ontological models, (2) extracting the corresponding time series data from the Building Automation System, and (3) performing computations and visualizations tailored to the user's query. The proposed method simplifies access to BAS data, allowing both domain experts and non-specialists to conduct sophisticated analyses without needing extensive technical knowledge of semantic web technologies. By demonstrating this pipeline, we facilitate more accessible and scalable data-driven decision-making in building operations and management.

Mulayim, Ozan Baris↗

Extraction and Analysis of Time Series Data from Building Automation Systems Using Large Language Models

Semantic schemas like Haystack 4, Brick and ASHRAE standard 223 enable the structured, standardized, and machine-readable representation of building data, facilitating interoperability, data integration, and advanced analytics. However, extracting information from these models requires specialized expertise in SPARQL and other programming languages, skills that are not commonly found among building professionals. Recent advancements in Large Language Models (LLMs), such as ChatGPT, enable the construction of queries using natural language, making it easier for individuals to interact with these systems in a manner that resembles everyday speech. However, these methods have not yet been tested on building semantic ontologies. This paper introduces a novel workflow and tool for enabling users to ask questions about a specific building's data, using natural language and receive answers automatically generated by GPT-4o. Our approach integrates semantic ontologies with advanced LLM capabilities to automate three critical steps: (1) generating SPARQL queries to retrieve time series references from ontological models, (2) extracting the corresponding time series data from the Building Automation System, and (3) performing computations and visualizations tailored to the user's query. The proposed method simplifies access to BAS data, allowing both domain experts and non-specialists to conduct sophisticated analyses without needing extensive technical knowledge of semantic web technologies. By demonstrating this pipeline, we facilitate more accessible and scalable data-driven decision-making in building operations and management.

Mulayim, Ozan Baris↗

American-Made Solar Prize: Edgeli Enables DER Integration (CRADA 615) (Final Report)

The purpose of this project was to demonstrate how granular time series data and automated data transformation, and impact assessment tools could speed interconnection approvals for distributed energy resource projects of various types and sizes. Types included community solar, rooftop solar, and EV charging projects. Using software routines to automate the transformation of data (e.g. GIS) to a network database and power flow model then applying scenarios to create hourly (8760) hosting capacity values and voltage and thermal impacts for specific projects, we were able to demonstrate the feasibility of quickly assembling and analyzing key utility data sets for interconnection purposes. The outcomes of this effort will become the foundation for future work that will enhance and encapsulate the software components developed as part of this project, into web services (e.g. APIs) that can be integrated into queue management systems and automate interconnection screening processes.

14 SOLAR ENERGY↗

Gauging nexus between topological and fracton phases

Coupled layer constructions are a valuable tool for capturing the universal properties of certain interacting quantum phases of matter in terms of the simpler data that characterizes the underlying layers. In the study of fracton phases, the X-Cube model in 3+1D can be realized via such a construction by starting with a stack of 2+1D Toric Codes and turning on a coupling which condenses a composite "particle-string" object. In a recent work [Phys. Rev. B 112, 125124 (2025)], we have demonstrated that in fact, the particle-string can be viewed as a symmetry defect of a topological 1-form symmetry. In this paper, we study the result of gauging this symmetry in depth. We unveil a rich gauging web relating the X-Cube model to symmetry protected topological (SPT) phases protected by a mix of subsystem and higher-form symmetries, subsystem symmetry fractionalization in the 3+1D Toric Code, and non-trivial extensions of topological symmetries by subsystem symmetries. Here, our work emphasizes the importance of topological symmetries in non-topological, geometric phases of matter.

Anyons↗

Open Power System Datasets and Open Simulation Engines: A Survey Toward Machine Learning Applications

A major factor behind the success of machine learning (ML) models in multiple domains is the availability and accessibility of large, labeled, and well-organized datasets for training and benchmarking. In comparison, power grid datasets face three major challenges: (i) real-world data is often restricted by regulatory constraints, privacy reasons, or security concerns, making it difficult to obtain and work with; (ii) synthetic datasets, which are created to address these limitations, often have incomplete information and are released using specialized tools, making them inaccessible to the broader community; and, (iii) input-output datasets are difficult to generate through simulation for non-experts because open-source simulators are not known outside the power system community. This survey addresses these challenges by serving as an entry point to publicly available datasets and simulators for researchers venturing in this area. We review the current landscape of open-source power network data, machine models, consumer demand profiles, renewable generation data, and inverter models. We also examine open-source power system simulators, which are crucial for generating high-quality, high-fidelity power grid datasets. We aim to provide a foundation for overcoming data scarcity and advance towards a structured web of datasets and simulators to support the development of ML for power systems.

42 ENGINEERING↗

Atacama Large Aperture Submillimeter Telescope (AtLAST) science: Resolving the hot and ionized Universe through the Sunyaev-Zeldovich effect

An omnipresent feature of the multi-phase “cosmic web” — the large-scale filamentary backbone of the Universe — is that warm/hot (≳ 10 5 K) ionized gas pervades it. This gas constitutes a relevant contribution to the overall universal matter budget across multiple scales, from the several tens of Mpc-scale intergalactic filaments, to the Mpc intracluster medium (ICM), all the way down to the circumgalactic medium (CGM) surrounding individual galaxies, on scales from ~ 1 kpc up to their respective virial radii (~ 100 kpc). The study of the hot baryonic component of cosmic matter density represents a powerful means for constraining the intertwined evolution of galactic populations and large-scale cosmological structures, for tracing the matter assembly in the Universe and its thermal history. To this end, the Sunyaev-Zeldovich (SZ) effect provides the ideal observational tool for measurements out to the beginnings of structure formation. The SZ effect is caused by the scattering of the photons from the cosmic microwave background off the hot electrons embedded within cosmic structures, and provides a redshift-independent perspective on the thermal and kinematic properties of the warm/hot gas. Still, current and next-generation (sub)millimeter facilities have been providing only a partial view of the SZ Universe due to any combination of: limited angular resolution, spectral coverage, field of view, spatial dynamic range, sensitivity, or all of the above. In this paper, we motivate the development of a wide-field, broad-band, multi-chroic continuum instrument for the Atacama Large Aperture Submillimeter Telescope (AtLAST) by identifying the scientific drivers that will deepen our understanding of the complex thermal evolution of cosmic structures. On a technical side, this will necessarily require efficient multi-wavelength mapping of the SZ signal with an unprecedented spatial dynamic range (from arcsecond to degree scales) and we employ detailed theoretical forecasts to determine the key instrumental constraints for achieving our goals.

79 ASTRONOMY AND ASTROPHYSICS↗

Monte Carlo Simulation with CAD Interface for Calculation of 3D Maps of Residual Dose (CRADA)

Objective: To develop an easy-to-use software application to predict and mitigate radiation effects in research environment, space instruments, nuclear plants and medical facilities and help nonproliferation and national security efforts. Tech-X will develop standalone software libraries and command-line tools for ( 1) translating CAD into tessellated surfaces and tetrahedral meshes in GDML (for Geant4 and MARS 15), ROOT (for MARS 15) and HDF5 (for compact representation and for the visualization) formats, (2) healing CAD geometries to make them suitable for Monte Carlo simulations; (3) creating uniform and variable Cartesian and cylindrical meshes for detailed scoring; and ( 4) efficient Monte Carlo navigation in CAD geometries. JLAB will finish automation of simulations of residual dose in CAD geometries and integrate Tech-X software into Geant4 and MARS15. Finally, Tech-X will develop a Graphical User Interface to set up and heal CAD geometries, create input files, run and visualize simulations for residual dose. This application will run on local desktops, local and remote clusters and supercomputers and will be made available through public clouds, such as Amazon Web Services.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Fermilab s Transition to Token Authentication

Fermilab is the first High Energy Physics institution to transition from X.509 user certificates to authentication tokens in production systems. All of the experiments that Fermilab hosts are now using JSON Web Token (JWT) access tokens in their grid jobs. Many software components have been either updated or created for this transition, and most of the software is available to others as open source. The tokens are defined using the WLCG Common JWT Profile. Token attributes for all the tokens are stored in the Fermilab FERRY system which generates the configuration for the CILogon token issuer. High security-value refresh tokens are stored in Hashicorp Vault configured by htvault-config, and JWT access tokens are requested by the htgettoken client through its integration with HTCondor. The Fermilab job submission system jobsub was redesigned to be a lightweight wrapper around HTCondor. For automated job submissions a managed tokens service was created to reduce duplication of effort and knowledge of how to securely keep tokens active. The existing Fermilab file transfer tool ifdh was updated to work seamlessly with tokens, as well as the Fermilab POMS (Production Operations Management System) which is used to manage automatic job submission and the RCDS (Rapid Code Distribution System) which is used to distribute analysis code via the CernVM FileSystem. The dCache storage system was reconfigured to accept tokens for authentication in place of X.509 proxy certificates. As some services and sites have not yet implemented token support, proxy certificates are still sent with jobs for backwards compatibility but some experiments are beginning to transition to stop using them. There have been some glitches and learning curve issues but in general the system has been performing well and is being improved as operational problems are addressed.

Dykstra, David↗

How can an ecosystem approach support integrated management of marine renewable energy? An initial assessment from an environmental point of view

With the increasing installation of marine renewable energy (MRE) devices in areas already subject to multiple anthropogenic activities and environmental changes, it is necessary to develop tools and methods for the integrated management of marine ecosystems. The ecosystem approach is a holistic environmental management method that considers all components of an ecosystem. The ecosystem approach has demonstrated utility in the application to various anthropogenic activities and is relevant for consideration within the context of MRE. Indeed, many of the effects observed on marine ecosystems from those other activities are also applicable to MRE development. This review is an initial assessment where we summarize the potential effects of MRE development on marine ecosystems and propose schematic frameworks for applying the ecosystem approach to MRE. We also provide a non-exhaustive list of commonly used models pertinent to the ecosystem approach and associated with several reference studies. An outline of core questions that can currently be answered using available modeling tools central to the ecosystem approach is provided, along with recommendations for the application of this approach to the MRE context. Further, we identify key knowledge gaps and areas that require additional investigation for meaningful application of the ecosystem approach to MRE development. Our recommendations mainly concern the current limitations of applying the ecosystem approach to concrete cases, such as consolidating knowledge of the effects of MRE on the local environment, the need to obtain fine-scale data, considering effects at different spatiotemporal scales, and, finally, the need for an interdisciplinary vision.

16 TIDAL AND WAVE POWER↗