Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “compression techniques”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

SymbolNet: neural symbolic regression with adaptive dynamic pruning for compression

Abstract Compact symbolic expressions have been shown to be more efficient than neural network (NN) models in terms of resource consumption and inference speed when implemented on custom hardware such as field-programmable gate arrays (FPGAs), while maintaining comparable accuracy (Tsoi et al 2024 EPJ Web Conf. 295 09036). These capabilities are highly valuable in environments with stringent computational resource constraints, such as high-energy physics experiments at the CERN Large Hadron Collider. However, finding compact expressions for high-dimensional datasets remains challenging due to the inherent limitations of genetic programming (GP), the search algorithm of most symbolic regression (SR) methods. Contrary to GP, the NN approach to SR offers scalability to high-dimensional inputs and leverages gradient methods for faster equation searching. Common ways of constraining expression complexity often involve multistage pruning with fine-tuning, which can result in significant performance loss. In this work, we propose S y m b o l N e t , a NN approach to SR specifically designed as a model compression technique, aimed at enabling low-latency inference for high-dimensional inputs on custom hardware such as FPGAs. This framework allows dynamic pruning of model weights, input features, and mathematical operators in a single training process, where both training loss and expression complexity are optimized simultaneously. We introduce a sparsity regularization term for each pruning type, which can adaptively adjust its strength, leading to convergence at a target sparsity ratio. Unlike most existing SR methods that struggle with datasets containing more than O ( 10 ) inputs, we demonstrate the effectiveness of our model on the LHC jet tagging task (16 inputs), MNIST (784 inputs), and SVHN (3072 inputs).

Tsoi, Ho Fung (ORCID:0000000225502184)↗

Neural architecture codesign for fast physics applications

We develop a pipeline to streamline neural architecture codesign for physics applications to reduce the need for ML expertise when designing models for novel tasks. Our method employs neural architecture search and network compression in a two-stage approach to discover hardware efficient models. This approach consists of a global search stage that explores a wide range of architectures while considering hardware constraints, followed by a local search stage that fine-tunes and compresses the most promising candidates. We exceed performance on various tasks and show further speedup through model compression techniques such as quantization-aware-training and neural network pruning. We synthesize the optimal models to high level synthesis code for FPGA deployment with the hls4ml library. Additionally, our hierarchical search space provides greater flexibility in optimization, which can easily extend to other tasks and domains. We demonstrate this with two case studies: Bragg peak finding in materials science and jet classification in high energy physics, achieving models with improved accuracy, smaller latencies, or reduced resource utilization relative to the baseline models.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Demonstration of a quantum-classical coprocessing protocol for simulating nuclear reactions

Quantum computers hold great promise for exact simulations of nuclear dynamical processes (e.g., scattering and reactions), which are paramount to the study of nuclear matter at the limit of stability and in the formation of chemical elements in stars. However, quantum simulations of the unitary (real) time dynamics of fermionic many-body systems require a currently prohibitive number of reliable and long-lived qubits. Here we propose a co-processing algorithm for the simulation of real-time dynamics in which the time evolution of the spatial coordinates is carried out on a classical processor, while the evolution of the spin degrees of freedom is carried out on quantum hardware. We demonstrate this hybrid scheme with the simulation of two neutrons scattering at the Lawrence Berkeley National Laboratory's Advanced Quantum Testbed. After implementing error mitigation strategies to improve the accuracy of the algorithm in addition to a combination of circuit compression techniques and tomography as methods to elucidate the onset of decoherence, our results validate the principle of the proposed co-processing scheme. A generalization of this present scheme will open the way for (real-time) path integral simulations of nuclear scattering.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Latent Space Dynamics Identification

LaSDI is a data-driven physical simulation software that forms a latent space for a given high-fidelity model and discovers a set of ordinary differential equations for the latent space dynamics. It allows a fast and accurate solution process, which is useful for multi-query decision making applications, such as design optimization and uncertainty quantification. The performance of the LaSDI framework is demonstrated on four different problems, i.e., 1D and 2D Burgers equations, nonlinear heat conduction, and radial advection problems. Both linear and nonlinear compression techniques, such as neural network and proper orthogonal decomposition, are used to form a latent space. A concept of local dynamics identification procedure is introduced to enable a parametric model, which enhances the accuracy level over a given parameter space.

Fries, William↗

Data-Driven Smoothers for Extreme-Scale Computing

Patch-based relaxation refers to a family of methods for solving linear systems which partitions the matrix into smaller pieces often corresponding to groups of adjacent degrees of freedom residing within patches of the computational domain. The two most common families of patch-based methods are block-Jacobi and Schwarz methods, where the former typically corresponds to non-overlapping domains and the later implies some overlap. We focus on cases where each patch consists of the degrees of freedom on a finite element method mesh cell. Patch methods often capture complex local physics much more effectively than simpler point-smoothers such as Jacobi; however, forming, inverting, and applying each patch can be prohibitively expensive in terms of both storage and computation time. To this end, we propose several approaches for performing analysis on these patches and constructing a reduced representation. The compression techniques rely on either matrix norm comparisons or unsupervised learning via a clustering approach. We illustrate how it is frequently possible to retain/factor less than 5% of all patches and still develop a method that converges only a little slower than when all patches are stored/factored.

97 MATHEMATICS AND COMPUTING↗

Oak Ridge National Laboratory Thermoplastic Composite Joining of Pultruded Continuous Fiber Spars for Multi-Functional Battery Enclosure

Oak Ridge National Laboratory (ORNL) worked with DowAksa Industry to investigate the potential of using PA66 material to produce larger composite parts, transcending the limitations of traditional presses. The project aimed to develop a suitable material for printing battery boxes using Nylon 66 infused with 40 wt.% carbon fiber, compatible with the ORNL Additive Manufacturing and Compression Molding (AM-CM) system. However, even with extensive work, it was concluded that the material is not compatible with the AMCM process and suitable for joining using compression technique. Main reason for the incompatibility were found to be very high crystallinity of the PA66 grade chosen for this work. To improve printability, a modifier (proprietary) was added to decrease crystallinity, but this led to reduced thermal stability and adhesion issues.

36 MATERIALS SCIENCE↗

Viscosity Measurements in Extreme Conditions

Radiation-hydrodynamics simulations are critical to the NNSA complex for understanding high energy density physics and for addressing problems in national security science. However, current rad-hydro codes typically do not account for viscosity, which generally leads to large uncertainties in problems involving mixing of materials under shock loading. This report details progress made towards experimental measurements of fluid viscosity at high pressures and temperatures achieved by shock wave compression techniques, as well as complementary material model development and hydrodynamic simulations.

36 MATERIALS SCIENCE↗

High-Peak-Power Long-Wave Infrared Lasers with CO2 Amplifiers

Long-wave infrared (LWIR) picosecond pulses with multi-terawatt peak power have recently become available for advanced high-energy physics and material research. Multi-joule pulse energy is achieved in an LWIR laser system via amplification of a microjoule seed pulse with high-pressure, mixed-isotope CO 2 amplifiers. A chirped-pulse amplification (CPA) scheme is employed in such a laser to reduce the nonlinear interaction between the optical field and the transmissive elements of the system. Presently, a research and development effort is underway towards an even higher LWIR peak power that is required, for instance, for promising particle acceleration schemes. The required boost of the peak power can be achieved by reducing the pulse duration to fractions of a picosecond. For this purpose, the possibility of reducing the gain narrowing in the laser amplifiers and post-compression techniques are being studied. Another direction in research is aimed at the increased throughput (i.e., repetition rate), efficiency, and reliability of LWIR laser systems. The transition from a traditional electric-discharge pumping to an optical pumping scheme for CO 2 amplifiers is expected to improve the robustness of high-peak-power LWIR lasers, making them suitable for broad implementation in scientific laboratory, industrial, and clinical environments.

43 PARTICLE ACCELERATORS↗

Technology requirements for post-1985 communications satellites

The technical and functional requirements for commercial communication satellites are discussed. The need for providing quality service at an acceptable cost is emphasized. Specialized services are postulated in a needs model which forecasts future demands. This needs model is based upon 322 separately identified needs for long distance communication. It is shown that the 1985 demand for satellite communication service for a domestic region such as the United States, and surrounding sea and air lanes, may require on the order of 100,000 MHz of bandwith. This level of demand can be met by means of the presently allocated bandwidths and developing some key technologies. Suggested improvements include: (1) improving antennas so that high speed switching will be possible; (2) development of solid state transponders for 12 GHz and possibly higher frequencies; (3) development of switched or steered beam antennas with 10 db or higher gain for aircraft; and (4) continued development of improved video channel compression techniques and hardware.

Burtt, J. E.↗

Solutions of the Navier-Stokes equations for vortex breakdown

Steady solutions of the Navier-Stokes equations, in terms of velocity and pressure, for breakdown in an unconfined viscous vortex are obtained numerically using the artificial compressibility technique of Chorin combined with an ADI finite-difference scheme. Axisymmetry is assumed and boundary conditions are carefully applied at the boundaries of a large finite region in an axial plane while resolution near the axis is maintained by a coordinate transformation. The solutions, which are obtained for Reynolds numbers up to 200 based on the free-stream axial velocity and a characteristic core radius, show that breakdown results from the diffusion and convection of vorticity away from the vortex core which, because of the strong coupling between the circumferential and axial velocity fields in strongly swirling flows, can lead to stagnation and reversal of the axial flow near the axis.

Grabowski, W. J.↗

Radar satellite altimetry and ocean wave height estimation

The design of a radar satellite altimeter having a plus or minus 10 cm topographic resolution at 20 meter (peak-to-trough) ocean wave heights is described. In addition to altimetry, the resulting design also provides a measurement of significant wave height over the range of 1.0 to 20 meters to within plus or minus 10%. A full deramp pulse compression technique followed by an analog filter bank to separate individual range returns is used in the radar transmitter/receiver design to reduce the A/D converter bandwidth from a rather impractical 330 MHz to less than 1 MHz. The altimeter design utilizes an onboard maximum likelihood estimate (MLE) processor to achieve the plus or minus 10 cm topographic resolution. It is shown that an MLE processor provides simultaneous optimum (minimum variance) estimates of satellite altitude, ocean wave height and electromagnetic ocean surface reflectivity.

Dooley, R. P.↗

Study of efficient video compression algorithms for space shuttle applications

Results are presented of a study on video data compression techniques applicable to space flight communication. This study is directed towards monochrome (black and white) picture communication with special emphasis on feasibility of hardware implementation. The primary factors for such a communication system in space flight application are: picture quality, system reliability, power comsumption, and hardware weight. In terms of hardware implementation, these are directly related to hardware complexity, effectiveness of the hardware algorithm, immunity of the source code to channel noise, and data transmission rate (or transmission bandwidth). A system is recommended, and its hardware requirement summarized. Simulations of the study were performed on the improved LIM video controller which is computer-controlled by the META-4 CPU.

Poo, Z.↗

An optimized buffer controlled data compression system

The digital data compression system considered uses a buffer controlled aperture algorithm which minimizes the mean-squared error between the reconstructed receiver output and transmitter input. The data compression technique selected is based on the zero-order floating aperture prediction rule. It is assumed that the statistics of the input data are initially uniformly distributed, stationary, and first-order Markov. The problem is solved for stationary data. An approach is presented for extending the results to slowly varying uniformly distributed nonstationary Markov data.

Dosik, P. H.↗