Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “control co-design”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Co-design of a wave energy converter through bi-conjugate impedance matching

As with other oscillatory power conversion systems, the design of wave energy converters can be understood as an impedance matching problem. By representing the wave energy converter as a multi-port network, two separate but related impedance matching conditions can be established. Satisfying these conditions maximizes power transfer to the load. In practice, these impedance matching conditions may be used to influence the design of the system (including the hull, power take-off, controller, mooring, etc.). To this end, this paper considers some example applications of wave energy converter design with the help of the impedance matching framework.

WEC↗

Optimal Co-Design of Integrated Thermal-Electrical Networks and Control Systems for Grid-interactive Efficient District (GED) Energy Systems

This project advances a unified, open-source framework for the optimal co-design of thermal, electrical, and control systems in grid-interactive efficient districts (GEDs). As communities integrate growing levels of distributed energy resources, traditional approaches that model thermal and electrical networks independently lead to reduced efficiency, limited flexibility, and missed opportunities for coordinated operation. To address these challenges, the research team developed a comprehensive suite of physics-based models, control algorithms, and software tools that enable holistic simulation, optimization, and demonstration of district-scale energy systems.

14 SOLAR ENERGY↗

CIRCLES: Congestion Impacts Reduction via CAV-in-the-loop Lagrangian Energy Smoothing

The energy efficiency of today’s vehicular mobility relies on the un-integrated combination of i) control via static assets (traffic lights, metering, variable speed limits, etc.); and ii) onboard vehicle automation (adaptive cruise control (ACC), ecodriving, etc.). These two families of control were not co-designed and are not engineered to work in coordination. Recent studies have shown i) limitations of controls, and even ii) negative impacts of ACC. This project focused on the technology development, implementation and prototyping, and validation of Mobile Traffic Control (MTC). MTC can be viewed as an extension of classical traffic control (in which static infrastructure actuates traffic flow). In the MTC paradigm, automated vehicles actuate the entire flow via their behavior, offering enhanced possibilities to optimize the energy footprint of traffic, if designed correctly. We set out to demonstrate for the first time that considerably reduced fuel consumption of all vehicles in traffic can be achieved via distributed control of a small proportion of CAVs. Compared to baseline vehicular technologies, our work offers a significant design departure: control algorithms for the CAVs consider the impact one vehicle can have on overall traffic, improving resulting overall fuel consumption. We focus on using a few vehicles (as traffic controllers via CAV technology) to improve the energy efficiency of traffic flow to further optimize energy efficiency. A live-traffic demonstration in November 2022 featured the deployment of 100 specially-equipped CAVs, constituting an approximate local penetration rate upwards of 2%.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

From IMT Device Measurements to Network-Level Consequences: When Learning Suppresses Beyond-LIF Neuron Dynamics

Emerging neuromorphic devices such as insulator--metal transition (IMT) devices exhibit complex temporal dynamics, including slow internal state memory, hysteresis, and burst-like firing, which are poorly captured by conventional leaky integrate-and-fire (LIF) neurons. However, it remains unclear when such dynamics influence learning and inference at the network level, particularly under commonly used unsupervised plasticity rules. We present a controlled, full-stack co-design study spanning experimental characterization of individual IMT devices, compact neuron model development, and large-scale spiking network simulations with identical architectures and learning rules. Rather than optimizing benchmark accuracy, our goal is to diagnose when neuron-level dynamics survive learning and competition, and when they are suppressed, to inform the co-design of devices, networks, and learning rules that can exploit beyond-LIF complexity.

42 ENGINEERING↗

Lessons learned from the design and operation of a small-scale cross-flow tidal turbine

In 2023, a first-generation prototype of a small-scale marine current turbine was operated in Sequim Bay, Washington (USA) for 141 days. The system, referred to as the Turbine Lander, was the product of a laboratory-to-field effort to develop a system that enables enhanced ocean sensing or vehicle recharge in remote, energetic settings. The turbine consists of a vertical-axis, cantilevered rotor (1.19 m x 0.85 m) with four foils installed on a gravity foundation. A broader range of constraints including the deployment strategy, site characteristics, and estimated loads, drove the system’s design. This work presents the design, characterization, operation, and post-recovery engineering assessment of the Turbine Lander. Pre-deployment characterization efforts yielded a peak power coefficient of approximately 0.3 for the rotor, although system losses resulted in much lower water-to-wire efficiencies under most operating conditions. The results demonstrate the importance of co-design among key components of the powertrain and control systems to achieve acceptable system efficiency across operating conditions.

Co-design↗

Next-generation electronics by co-design with chalcogenide materials

As Moore’s law approaches its limits, chalcogenides offer a promising route to next-generation computing and sensing, thanks to their topological, magnetoelectric, excitonic, and spintronic properties. Yet the same traits that make them appealing, e.g., exotic properties at monolayer thickness, clean van der Waals interfaces, and strong many-body effects, also heighten sensitivity to fabrication, hindering translation into scalable devices. Progress is further constrained by fragmented knowledge across synthesis, processing, and integration, and by the lack of systematic links between fabrication parameters and performance metrics. This Perspective examines key obstacles in controlling chalcogenide heterostructures and stresses the need for an integrated co-design framework, uniting materials growth, processing, and device architecture to accelerate practical technologies.

36 MATERIALS SCIENCE↗

Co-Design of Charge Transport Superhighways to Connect Catalytic Sites in Soft Photoelectrochemical Systems

Efficient photon-to-electron-to-molecule conversion requires multi-length scale control over charge transport pathways, where electronic charges are delivered to catalytic sites under high mass transport flux. A fundamental question is how can we co-design charge transport pathways to promote efficient charge transfer to/from catalytic sites in complex three-dimensional architectures? Soft conducting polymer systems offer exceptional promise to provide three-dimensional charge transport networks, where electrolyte (ion and solvent) can interdiffuse to promote long-lived charge carriers and the molecular nature allows for strategic synthetic design of catalytic sites. Herein we combine theoretical and experimental approaches to investigate the earliest stages of photoelectrochemical deposition of near-surface catalytic sites (Pt) on soft bulk heterojunction polymeric semiconductors composed of a prototype donor (PTB7-Th) and a prototype acceptor (N2200) as a model system towards better understanding molecular catalyst-polymer site interactions. We focus initially on photoelectrochemical deposition of low Pt loadings, nanoparticle sizes (formed by progressive nucleation) below 20 nm, for both density functional theory (DFT) modeling studies and for spectroscopic characterization using surface-sensitive X-ray and UV-photoemission (XPS/UPS). DFT modeling of “n-type” N2200 slabs reveal for the first time that sulfur atoms in the thiophene units serve as the lowest-energy adsorption sites for single Pt atoms, while larger Pt clusters engage more complexly with both thiophene and naphthalene diimide (NDI) core sites. Changes in chemical composition observed by X-ray photoelectron spectroscopy (XPS) support the DFT predictions, and the angle-resolved measurements reveal that Pt nucleation initiates at subsurface sites which appear to be localized active domains that promote charge transport/transfer and enable vertical growth toward the surface. These results suggest that light-activated Pt nanoparticle deposition decorates energetically distinct sites, where photoactivity is dictated by the local energetics of those sites, and the fact that they represent the termini of charge transport “super-highways” – a small percentage of the total volume of the donor/acceptor polymeric active layer which carries most of the photocurrent generated during both Pt deposition and photoelectrochemical HER. We posit that these initial studies provide a foundational strategy for design of catalytic sites in the near surface regions of complex polymeric materials and advancing soft semiconductor-based photoelectrochemical systems. Achieving a nanometer-scale understanding of catalyst deposition and the impact of local composition and energetics on that placement, should ultimately provide the design guidelines (co-design) for a broad array of catalysts at sites that optimize that efficiency and maximize platform durability.

14 SOLAR ENERGY↗

Optimization-based approaches to control of connected and automated vehicles: Principles, complexities, applications, challenges, and outlook

Safe and optimal motion control for connected and automated vehicles (CAVs) poses a fundamental optimization challenge at the intersection of system complexity, environmental uncertainty, and stringent real-time constraints. Existing surveys address this challenge in isolation – focusing either on specific control techniques or individual uncertainty sources – without providing a unified framework that characterizes the trade-offs among computational tractability, performance verifiability, and adaptive generalization across paradigms. This review addresses that gap by presenting a cohesive analytical framework concentrated on the decision-making and trajectory optimization layers of the CAV autonomy stack. We systematically analyze three major optimization paradigms – first-principles model-based optimization, data-driven methods, and hybrid synergistic architectures – evaluating each against four core complexity axes: problem formulation, constraint handling, optimality guarantees, and robustness. Key applications including platooning, trajectory planning, collision avoidance, and cooperative control are examined to reveal recurring methodological patterns and critical operational constraints that limit real-world performance. Our synthesis identifies verifiable hybrid architectures, incentive-aligned multi-agent cooperation, and hardware-algorithm co-design as the defining research frontiers, and distills a targeted agenda for developing CAV control systems that are simultaneously safe, computationally efficient, and deployable in the full complexity of real-world traffic environments.

Muzahid, Abu Jafar Md [University of Tennessee, Kn↗

Protection of Inverter-Dependent Transmission Systems (PROTECT-IT)

This presentation highlights the overall objectives of the SETO funded protection project. The main technical approaches are also highlighted. This high impact project produces multiple innovative outcomes, including comprehensive impact study of how IBR affects protection elements, simplified low-order IBR model for protection engineers, enhanced and data-driven protection design, co-design concept for coordination protection and IBRs.

14 SOLAR ENERGY↗

Intelligent Experiments through Real-Time AI: Fast Data Processing and Autonomous Detector Control for High-Energy Nuclear Experiments

The aim of this project is to develop software and hardware for fast real-time data processing and autonomous detector control and calibration for the sPHENIX and the future EIC experiments. Below summarizes Georgia Tech team efforts in the past year: 1. We developed a real-time clustering algorithm and FPGA-based pipeline architecture for processing fired pixel data from ALPIDE sensors in sPHENIX experiments. Our Columnar Clustering Co-Design introduces a hardware-aware, stream-friendly approach that segments pixel data by column pairs using a Column Pair Clustering (CPC) strategy, followed by Cluster Stitching to merge adjacent subclusters. Implemented in Vitis HLS, the pipeline comprises five stages—read-in, subclustering, stitching, analysis, and write-out—connected by tagged HLS streams with custom end-of-event signaling for robust synchronization. We designed a pipelined dataflow model optimized for throughput, low latency, and minimal buffering, enabling scalable clustering across events of arbitrary size. Our system maintains spatial precision via center-of-mass and shape key extraction and efficiently handles edge cases such as fragmented or nested clusters. Compared against DBSCAN in both software and hardware, our approach demonstrates competitive performance under FPGA constraints. 2. We also conducted a comprehensive algorithm-to-hardware co-design of connected component analysis tailored for sPHENIX experiments, focusing on real-time, low-latency processing using FPGAs and High-Level Synthesis (HLS). Starting from a Python-based particle tracking pipeline, the team translated the core logic—graph traversal via DFS and Union-Find—into an HLS-compatible C++ model, replacing dynamic memory and recursion with static arrays and pipelined control flow. The final design includes a fully streamed and dataflow-compatible Union-Find kernel optimized across five iterations, incorporating loop pipelining, array partitioning, AXI/FIFO interface tuning, and function flattening. Experimental results show up to 14.8× speedup over the CPU baseline, reducing per-graph latency to 1.58 μs and demonstrating strong resource efficiency with only ~7k LUTs and zero BRAM usage. The design maintains functional correctness against the Python reference using a Python-based C-simulation framework and Mean Squared Error metrics. This work validates the potential of HLS-driven FPGA designs for edge-level HEP data acquisition, laying a scalable foundation for future integration with real-time detector pipelines and multi-graph processing systems.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

An Integrated Framework for Memory-Centric Analysis: From Trace Collection to Co-Design

The memory wall phenomenon—where advances in processor performance significantly outpace those in memory subsystems—poses a fundamental challenge for contemporary computing systems. In memory-bound applications, memory subsystem behavior dominates performance, yet existing analysis approaches present significant limitations: detailed microarchitectural simulators require days to weeks to simulate modest workloads; hardware performance counters provide only aggregate statistics that obscure temporal and spatial access patterns; and scaled simulation approaches face challenges in capturing certain behaviors that emerge at larger scales. These limitations reflect a processor-centric design philosophy increasingly misaligned with memory-bound workloads where detailed understanding of memory access patterns, cache hierarchy interactions, and contention is critical for effective optimization. This paper presents an integrated framework for memory-centric analysis that enables effective hardware-software co-design. We describe practical trace collection techniques, including hardware-assisted processor tracing with minimal overhead and portable software-based instrumentation with statistical sampling. We present multi-perspective analysis methods that examine memory behavior from temporal, sequential, spatial, and relational viewpoints, revealing distinct optimization opportunities invisible in aggregate metrics. We detail an architectural modeling framework that uses sampled traces with temporal interpolation and confidence-based filtering to evaluate cache and memory configurations. Evaluation on representative benchmarks demonstrates that this framework achieves practical accuracy (L2 cache errors of 2.64\%, confidence-filtered L3 errors of 9.92\%, bandwidth errors of 7.33\%) while providing substantial speedup (26.8×) over cycle-accurate simulation, enabling rapid design space exploration. We demonstrate how this integrated framework enables systematic identification of both hardware optimizations (memory controller tuning, bank partitioning, NUMA configuration) and software optimizations (data layout restructuring, prefetching strategies, memory-aware scheduling). Through this comprehensive treatment of the memory-centric analysis pipeline—from trace collection through architectural modeling to co-design application—we provide researchers and practitioners with practical techniques for addressing memory bottlenecks in contemporary computing systems.

Gajaria, Dhruv Mayur↗

Cyber-Physical System: Design for Sustainability and Resilience

When considering the design tools needed in the transition from numeric models to pilot plant, cyber-physical systems (CPS) come to the forefront as a method to model complex integrated energy systems. CPS approach has proven to be valuable to identify opportunities for economically viable early adoption of integrated energy technologies. This tutorial will introduce the concepts and the roles of CPS in co-design to minimize risks for pilot plant and technology deployment. This tutorial will also layout basic requirements for the CPS development, which requires a highly interdisciplinary effort with expertise in sensors, hardware testing, real-time modeling, controls, and system integration.

Harun, Nor Farida↗

NeuroCoreX: An Open-Source FPGA-Based Spiking Neural Network Emulator with On-Chip Learning

Spiking Neural Networks (SNNs) are computational models inspired by the event-driven communication and connectivity patterns of biological neural circuits. They enable high energy efficiency and natural support for diverse architectures ranging from layered networks to small-world and graphstructured topologies. In this work, we introduce NeuroCoreX, an open-source, FPGA-based spiking neural network emulator that provides real-time, on-chip learning and flexible network organization. NeuroCoreX supports both feedforward sensory inputs streamed directly from sensors or PCs via UART and recurrent on-chip connectivity, enabling simultaneous processing and learning from external stimuli and internal network dynamics-capabilities rarely available in existing FPGA SNN platforms. The system implements a Leaky Integrate-and-Fire (LIF) neuron model with current-based synapses and supports pair-based STDP learning on both feedforward and recurrent synapses. A lightweight Python interface enables interactive configuration, live monitoring, weight read-back, and experiment control. Importantly, NeuroCoreX is tightly integrated with the SuperNeuroMAT simulator, allowing SNN models to be transferred seamlessly from software to hardware for hardware-in-the-loop development. By combining real-time plasticity, flexible connectivity, and an open-source VHDL implementation, NeuroCoreX provides an extensible and accessible platform for neuromorphic research, algorithm-hardware co-design, and energy-efficient edge intelligence.

Gautam, Ashish [ORNL]↗

InterQnet: A Heterogeneous Full-Stack Approach to Co-Designing Scalable Quantum Networks

Quantum communications have progressed significantly, moving from a theoretical concept to small-scale experiments to recent metropolitan-scale demonstrations. As the technology matures, it is expected to revolutionize quantum computing in much the same way that classical networks revolutionized classical computing. Quantum communications will also enable breakthroughs in quantum sensing, metrology, and other areas. However, scalability has emerged as a major challenge, particularly in terms of the number and heterogeneity of nodes, the distances between nodes, the diversity of applications, and the scale of user demand. This article describes InterQnet, a multidisciplinary project that advances scalable quantum communications through a comprehensive approach that improves devices, error handling, and network architecture. InterQnet has a two-pronged strategy to address scalability challenges: InterQnet-Achieve focuses on practical realizations of heterogeneous quantum networks by building and then integrating first-generation quantum repeaters with error mitigation schemes and centralized automated network control systems. The resulting system will enable quantum communications between two heterogeneous quantum platforms through a third type of platform operating as a repeater node. InterQnet-Scale focuses on a systems study of architectural choices for scalable quantum networks by developing forward-looking models of quantum network devices, advanced error correction schemes, and entanglement protocols. Here, we report our current progress toward achieving our scalability goals.

Chung, Joaquin [Argonne] (ORCID:0000000173833810)↗

SAN-Based Block Polymers as a Platform for Manufacturing Strong Isoporous Membranes

Ultrafiltration (UF) membranes are ubiquitous in water purification and bioprocessing. However, co-designing their mechanical and transport properties remains challenging because of the broad pore size distributions at the surface and within the bulk that result from nonsolvent-induced phase separation (NIPS) – their typical manufacturing process. These distributions influence the hydrodynamic resistance to water flow and the stress concentrations around the pores. Developing advanced UF membranes requires innovative molecular designs that offer control over the surface and bulk pores, as well as the mechanical properties of the load-bearing, polymer. Here, we introduce a platform for designing UF membranes by leveraging solution self-assembly of block polymers and chain architectures with pendant polar groups. The block polymers consist of a poly(styrene-co-acrylonitrile) hydrophobic block, which is known for its strength, and a poly(4-vinyl pyridine) hydrophilic block, which drives solution self-assembly. We focus on a series of block polymers with constant molecular weight, M n ≈ 115 kDa, SAN fraction, 75 wt.%, and varying acrylonitrile content, 0 to 40 mol%, to demonstrate that: (i) RAFT dispersion copolymerization of acrylonitrile and styrene provides a facile route to synthesize strong block polymers, (ii) incorporation of acrylonitrile into the hydrophobic block enhances membrane strength by facilitating chain entanglements and dipole-dipole interactions, and (iii) acrylonitrile alters the balance between membrane permeance and rejection, even when the membranes feature similar surface and bulk pores. Overall, our results provide insights into the molecular design of UF membranes with enhanced mechanical and separation properties, contributing to the development of materials for water and energy technologies.

deformation↗

Financial-technical co-design for capital-intensive, resource-responsive energy systems

Because of their capital-intensive operation, wind energy systems that are competitive in terms of the cost of the energy that they produce lead to risk-reward trade-offs that make their business cases less favorable than those of conventional energy generation technologies. However, wind energy systems tend to be designed to maximize energy production or minimize cost of energy rather than to maximize their business cases. In this work, we attempt to exploit designs specifically tailored to business cases. We develop a novel framework for analyzing energy systems that ties their design variables to monthly operating incomes using simple models and historical hourly market and resource data. Using this approach, we demonstrate that for a wind site with abundant wind resource in the California Independent System Operator market, we can control the trade-off between mean and 5th percentile monthly returns by choosing the specific power of the turbine at a fixed modeled initial capital cost. Our framework gives a measure of the risk-reward spectrum of energy generation assets that could be built at a given site with respect to the sub-annual resource/market variation.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

The Case for Co-Designing Model Architectures with Hardware

While GPUs are responsible for training the vast majority of state-of-the-art deep learning models, the implications of their architecture are often overlooked when designing new deep learning (DL) models. As a consequence, modifying a DL model to be more amenable to the target hardware can significantly improve the runtime performance of DL training and inference. In this paper, we provide a set of guidelines for users to maximize the runtime performance of their transformer models. These guidelines have been created by carefully considering the impact of various model hyperparameters controlling model shape on the efficiency of the underlying computation kernels executed on the GPU. We find the throughput of models with “efficient” model shapes is up to 39% higher while preserving accuracy compared to models with a similar number of parameters but with unoptimized shapes.

Yin, Junqi↗

QUCODE: End-to-End Qubit Co-Design

The design of a quantum computer can be broken down into different steps, e.g., the material science aspect of designing qubits and devices, considerations of controlling the state of the qubits and their environment, the computer science aspects of mapping algorithms to the available primitives of the quantum computer, and the programming of an application in terms of the available algorithms. Research in these areas is currently fairly isolated, and there is framework for an end-to-end design approach where a desired application informs the choice of materials for the qubits and their environment, and vice versa.We identify knowledge gaps and opportunities for research that builds on existing PNNL capabilities.

36 MATERIALS SCIENCE↗