Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Field programmable gate array”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

142 records · Page 8

Bridging Python to Silicon: The SODA Toolchain

Systems performing scientific computing, data analysis, and machine learning tasks have a growing demand for application-specific accelerators that can provide high computational performance while meeting strict size and power requirements. However, the algorithms and applications that need to be accelerated are evolving at a rate that is incompatible with manual design processes based on hardware description languages. Agile hardware design tools based on compiler techniques can help by quickly producing an application-specific integrated circuit (ASIC) accelerator starting from a high-level algorithmic description. Here, we present the software-defined accelerator (SODA) synthesizer, a modular and open-source hardware compiler that provides automated end-to-end synthesis from high-level software frameworks to ASIC implementation, relying on multilevel representations to progressively lower and optimize the input code. Our approach does not require the application developer to write any register-transfer level code, and it is able to reach up to 364 giga floating point operations per second (GFLOPS)/W efficiency (32-bit precision) on typical convolutional neural network operators.

97 MATHEMATICS AND COMPUTING↗

A Quench Detection and Monitoring System for Superconducting Magnets at Fermilab

A quench detection system was developed for protecting and monitoring the superconducting solenoids for the Muon-to-Electron Conversion Experiment (Mu2e) at Fermilab. The quench system was designed for a high level of dependability and long-term continuous operation. It is based on three tiers: Tier-I, FPGA-based Digital Quench Detection (DQD); Tier-II, Analog Quench Detection (AQD); and Tier-3, the quench controls and data management system. The Tier-I and Tier-II are completely independent and fully redundant systems. The Tier-3 system is based on National Instruments (NI) C-RIO and provides the user interface for quench controls and data management. It is independent from Tiers I & II. The DQD provides both quench detection and quench characterization (monitoring) capability. Both DQD and AQD have built-in high voltage isolation and user programmable gains and attenuations. The DQD and AQD also includes user configured current dependent thresholding and validation times. A 1 st article of the three-tier system was fully implemented on the new Fermilab magnet test stand for the HL-LHC Accelerator Up-grade Project (AUP). It successfully provided quench protection and monitoring (QPM) for a cold superconducting bus test in November 2020. The Mu2e quench detection design has since been implemented for production testing of the AUP magnets. A detailed description of the system along with results from the AUP superconducting bus test will be presented.

monitoring↗

Information Leakage Analysis using a Co-design-Based Fault Injection Technique on a RISC-V Microprocessor

The RISC-V instruction set architecture open licensing policy has spawned a hive of development activity, making a range of implementations publicly available. The environments in which RISC-V operates have expanded correspondingly, driving the need for a generalized approach to evaluating the reliability of RISC-V implementations under adverse operating conditions or after normal wear-out periods. Fault injection (FI) refers to the process of changing the state of registers or wires, either permanently or momentarily, and then observing execution behavior. The analysis provides insight into the development of countermeasures that protect against the leakage or corruption of sensitive information which might occur because of unexpected execution behavior. In this paper, we develop a hardware-software co-design architecture that enables fast, configurable fault emulation and utilize it for information leakage and data corruption analysis. Modern System-on-chip FPGAs enable building an evaluation platform where control elements run on a processor(s) (PS) simultaneously with the target design running in the programmable logic (PL). Software components of the FI system introduce faults and report execution behavior. A pair of RISC-V FI-instrumented implementations are created and configured to execute the Advanced Encryption Standard and Twister algorithms. Key and plaintext information leakage and degraded pseudo-random sequences are both observed in the output for a subset of the emulated faults.

42 ENGINEERING↗

Control System of Multi-Port Autonomous Reconfigurable Solar Power Plant (MARS) & HIL Platforms for Design

Multi-port autonomous reconfigurable solar power plant (MARS) provides an attractive alternative to connect photovoltaic (PV) and energy storage systems (ESSs) to high-voltage direct current (HVdc) links and high-voltage alternating current (ac) grids. In this paper, a unique hierarchical control system of MARS is proposed and evaluated. To evaluate the control system and associated algorithms in early-stage research of complex architectures like MARS, it is important to develop unique suitable hardware-in-the-loop (HIL) platforms. In this paper, the HIL platforms for MARS to evaluate the performance of the hierarchical control system and the control algorithms implemented are presented. Further, they help with the design process of control systems. The real-time simulation models and algorithms that are utilized for MARS in the HIL platforms are also discussed in the paper. The HIL experiments of the control system of MARS showcase the capability to provide continuity of operation under faults and frequency support to the power grid during loss of generation. They also showcase the stability of the proposed hierarchical control system of MARS.

14 SOLAR ENERGY↗

High-count rate effects in event processing for XRISM/ Resolve X-ray microcalorimeter: I. Ground test

The spectroscopic performance of an X-ray microcalorimeter is compromised at high count rates. We utilize the Resolve X-ray microcalorimeter onboard the XRISM satellite to examine the effects observed during high-count rate measurements and propose modeling approaches to mitigate them. We specifically address the following instrumental effects that impact performance: CPU limit, pile-up, and untriggered electrical cross-talk. Experimental data at high count rates were acquired during ground testing using the flight model instrument and a calibration X-ray source. In the experiment, data processing not limited by the performance of the onboard CPU was run in parallel, which cannot be done in orbit. This makes it possible to access the data degradation caused by limited CPU performance. We use these data to develop models that allow for a more accurate estimation of the aforementioned effects. To illustrate the application of these models in observation planning, we present a simulated observation of GX 13+1. Understanding and addressing these issues is crucial to enhancing the reliability and precision of X-ray spectroscopy in situations characterized by elevated count rates.

47 OTHER INSTRUMENTATION↗

Passband Signal Detection at the Edge

Algorithms for radio frequency (RF) spectrum awareness need to be compatible with edge hardware to be practical for many applications. We developed a signal detection and classification model for the ZCU111 RF System-on-a-Chip (RFSoC) that operates on the fast Fourier transform of passband RF data. The system can detect and classify multiple signals of interest and display the predictions in real-time. The model consists of a modified ConvNeXt backbone and YOLOv3 head to operate on the Deep Learning Processing Unit on the RFSoC. We gathered datasets for training and testing by using a software defined radio to transmit example signals of Wi-Fi 802.11 b/g, Wi-Fi 802.11 n, FM Radio, LTE and LTE-M. By leveraging multiple inputs on the RFSoC frontend, the datasets span up to 4 GHz of bandwidth. The models showed high performance in classification accuracy, center frequency error, bandwidth error, and detection accuracy for both single and multi-signal datasets.

42 ENGINEERING↗

Distance-Weighted Graph Neural Networks on FPGAs for Real-Time Particle Reconstruction in High Energy Physics

Graph neural networks have been shown to achieve excellent performance for several crucial tasks in particle physics, such as charged particle tracking, jet tagging, and clustering. An important domain for the application of these networks is the FGPA-based first layer of real-time data filtering at the CERN Large Hadron Collider, which has strict latency and resource constraints. We discuss how to design distance-weighted graph networks that can be executed with a latency of less than one μs on an FPGA. To do so, we consider a representative task associated to particle reconstruction and identification in a next-generation calorimeter operating at a particle collider. We use a graph network architecture developed for such purposes, and apply additional simplifications to match the computing constraints of Level-1 trigger systems, including weight quantization. Using the hls4ml library, we convert the compressed models into firmware to be implemented on an FPGA. Performance of the synthesized models is presented both in terms of inference accuracy and resource usage.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Experimental Validation of a Modular All-Electric Power Take-Off Topology for Wave Energy Converter Enabling Marine Renewable Energy Interconnection

Power electronic converters are an enabling technology for the emerging marine energy applications, such as using ocean waves to produce electricity. This paper outlines the power take-off system and its key components used in a wave energy converter offering modularity and scalability to generate power efficiently. The proposed power take-off system was implemented based on a modular multilevel converter and could be deployed to convert any alternating current electrical energy to a different alternating current for interconnection to grid or non-grid applications. Examples of widespread deployment are supplying electricity to coastal communities or producing clean drinking water. The analysis using both the simulation tests and laboratory experiments verified the design objectives and basic functionality of the developed power take-off system. An acceptable response using a field programmable gate array-based controlled laboratory testbench was achieved, complying with guidelines specified in the prevalent industry standards. Seamless operation during steady-state and transients for the studied wave energy converter was achieved as supported by the obtained results. The key findings of this work were experimentally examined under different load conditions, direct current bus voltage fluctuations, and generator speed–torque regulation. The ability of the power take-off system to generate high-power quality of the waveforms, e.g., against adhering to the IEEE 519-2022 standard for total harmonic distortion limits, is also confirmed.

Engineering↗

Field programmable spin arrays for scalable quantum repeaters

The large scale control over thousands of quantum emitters desired by quantum network technology is limited by the power consumption and cross-talk inherent in current microwave techniques. Here we propose a quantum repeater architecture based on densely-packed diamond color centers (CCs) in a programmable electrode array, with quantum gates driven by electric or strain fields. This ‘field programmable spin array’ (FPSA) enables high-speed spin control of individual CCs with low cross-talk and power dissipation. Integrated in a slow-light waveguide for efficient optical coupling, the FPSA serves as a quantum interface for optically-mediated entanglement. We evaluate the performance of the FPSA architecture in comparison to a routing-tree design and show an increased entanglement generation rate scaling into the thousand-qubit regime. Our results enable high fidelity control of dense quantum emitter arrays for scalable networking.

42 ENGINEERING↗

Compiling Quantum Circuits for Dynamically Field-Programmable Neutral Atoms Array Processors

Dynamically field-programmable qubit arrays (DPQA) have recently emerged as a promising platform for quantum information processing. In DPQA, atomic qubits are selectively loaded into arrays of optical traps that can be reconfigured during the computation itself. Leveraging qubit transport and parallel, entangling quantum operations, different pairs of qubits, even those initially far away, can be entangled at different stages of the quantum program execution. Such reconfigurability and non-local connectivity present new challenges for compilation, especially in the layout synthesis step which places and routes the qubits and schedules the gates. In this paper, we consider a DPQA architecture that contains multiple arrays and supports 2D array movements, representing cutting-edge experimental platforms. Within this architecture, we discretize the state space and formulate layout synthesis as a satisfiability modulo theories problem, which can be solved by existing solvers optimally in terms of circuit depth. For a set of benchmark circuits generated by random graphs with complex connectivities, our compiler OLSQ-DPQA reduces the number of two-qubit entangling gates on small problem instances by 1.7x compared to optimal compilation results on a fixed planar architecture. To further improve scalability and practicality of the method, we introduce a greedy heuristic inspired by the iterative peeling approach in classical integrated circuit routing. Using a hybrid approach that combined the greedy and optimal methods, we demonstrate that our DPQA-based compiled circuits feature reduced scaling overhead compared to a grid fixed architecture, resulting in 5.1X less two-qubit gates for 90 qubit quantum circuits. These methods enable programmable, complex quantum circuits with neutral atom quantum computers, as well as informing both future compilers and future hardware choices.

Physics↗

Beam steering at the nanosecond time scale with an atomically thin reflector

Techniques to mold the flow of light on subwavelength scales enable fundamentally new optical systems and device applications. The realization of programmable, active optical systems with fast, tunable components is among the outstanding challenges in the field. Here, we experimentally demonstrate a few-pixel beam steering device based on electrostatic gate control of excitons in an atomically thin semiconductor with strong light-matter interactions. By combining the high reflectivity of a MoSe 2 monolayer with a graphene split-gate geometry, we shape the wavefront phase profile to achieve continuously tunable beam deflection with a range of 10°, two-dimensional beam steering, and switching times down to 1.6 nanoseconds. Our approach opens the door for a new class of atomically thin optical systems, such as rapidly switchable beam arrays and quantum metasurfaces operating at their fundamental thickness limit.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Long-lived Bell states in an array of optical clock qubits

The generation of long-lived entanglement in optical atomic clocks is one of the main goals of quantum metrology. Arrays of neutral atoms, where Rydberg-based interactions may generate entanglement between individually controlled and resolved atoms, constitute a promising quantum platform to achieve this. Here we leverage the programmable state preparation afforded by optical tweezers and the efficient strong confinement of a three-dimensional optical lattice to prepare an ensemble of strontium-atom pairs in their motional ground state. We engineer global single-qubit gates on the optical clock transition and two-qubit entangling gates via adiabatic Rydberg dressing, enabling the generation of Bell states with a state-preparation-and-measurement-corrected fidelity of 92.8(2.0)% (87.1(1.6)% without state-preparation-and-measurement correction). For use in quantum metrology, it is furthermore critical that the resulting entanglement be long lived; we find that the coherence of the Bell state has a lifetime of 4.2(6) s via parity correlations and simultaneous comparisons between entangled and unentangled ensembles. Such long-lived Bell states can be useful for enhancing metrological stability and bandwidth. In the future, atomic rearrangement will enable the implementation of many-qubit gates and cluster state generation, as well as explorations of the transverse field Ising model.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Long-lived Bell states in an array of optical clock qubits

The generation of long-lived entanglement on an optical clock transition is a key requirement to unlocking the promise of quantum metrology. Arrays of neutral atoms constitute a capable quantum platform for accessing such physics, where Rydberg-based interactions may generate entanglement between individually controlled and resolved atoms. To this end, we leverage the programmable state preparation afforded by optical tweezers along with the efficient strong confinement of a 3d optical lattice to prepare an ensemble of strontium atom pairs in their motional ground state. We engineer global single-qubit gates on the optical clock transition and two-qubit entangling gates via adiabatic Rydberg dressing, enabling the generation of Bell states, $|\psi{\rangle}$ $= \frac{1}{\sqrt{2}}(|gg\rangle$ $+ i|ee\rangle{)}$, with a fidelity of $\mathcal{F}$ = 92.8(2:0)%. For use in quantum metrology, it is furthermore critical that the resulting entanglement be long lived; we find that the coherence of the Bell state has a lifetime of $τ_{bc}$ = 4:2(6) s via parity correlations and simultaneous comparisons between entangled and unentangled ensembles. Such Bell states can be useful for enhancing metrological stability and bandwidth. Further rearrangement of hundreds of atoms into arbitrary configurations using optical tweezers will enable implementation of many-qubit gates and cluster state generation, as well as explorations of the transverse field Ising model and Hubbard models with entangled or finite-range-interacting tunnellers.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗