Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “compressing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Impact of lossy compression on diagnostic accuracy of radiographs for periapical lesions

OBJECTIVES: The purpose of this study was to evaluate the lossy Joint Photographic Experts Group compression for endodontic pretreatment digital radiographs. STUDY DESIGN: Fifty clinical charge-coupled device-based, digital radiographs depicting periapical areas were selected. Each image was compressed at 2, 4, 8, 16, 32, 48, and 64 compression ratios. One root per image was marked for examination. Images were randomized and viewed by four clinical observers under standardized viewing conditions. Each observer read the image set three times, with at least two weeks between each reading. Three pre-selected sites per image (mesial, distal, apical) were scored on a five-scale score confidence scale. A panel of three examiners scored the uncompressed images, with a consensus score for each site. The consensus score was used as the baseline for assessing the impact of lossy compression on the diagnostic values of images. The mean absolute error between consensus and observer scores was computed for each observer, site, and reading session. RESULTS: Balanced one-way analysis of variance for all observers indicated that for compression ratios 48 and 64, there was significant difference between mean absolute error of uncompressed and compressed images (P <.05). After converting the five-scale score to two-level diagnostic values, the diagnostic accuracy was strongly correlated (R (2) = 0.91) with the compression ratio. CONCLUSION: The results of this study suggest that high compression ratios can have a severe impact on the diagnostic quality of the digital radiographs for detection of periapical lesions.

NASA Discipline Space Human Factors↗

Lumbar spine disc height and curvature responses to an axial load generated by a compression device compatible with magnetic resonance imaging

STUDY DESIGN: Axial load-dependent changes in the lumbar spine of supine healthy volunteers were examined using a compression device compatible with magnetic resonance imaging. OBJECTIVE: To test two hypotheses: Axial loading of 50% body weight from shoulder to feet in supine posture 1) simulates the upright lumbar spine alignment and 2) decreases disc height significantly. SUMMARY OF BACKGROUND DATA: Axial compression on the lumbar spine has significantly narrowed the lumbar dural sac in patients with sciatica, neurogenic claudication or both. METHODS: Using a device compatible with magnetic resonance imaging, the lumbar spine of eight young volunteers, ages 22 to 36 years, was axially compressed with a force equivalent to 50% of body weight, approximating the normal load on the lumbar spine in upright posture. Sagittal lumbar magnetic resonance imaging was performed to measure intervertebral angle and disc height before and during compression. RESULTS: Each intervertebral angle before and during compression was as follows: T12-L1 (-0.8 degrees +/- 2.5 degrees and -1.5 degrees +/- 2.6 degrees ), L1-L2 (0.7 degrees +/- 1.4 degrees and 3.3 degrees +/- 2.9 degrees ), L2-L3 (4.7 degrees +/- 3.5 degrees and 7.3 degrees +/- 6 degrees ), L3-L4 (7.9 degrees +/- 2.4 degrees and 11.1 degrees +/- 4.6 degrees ), L4-L5 (14.3 degrees +/- 3.3 degrees and 14.9 degrees +/- 1.7 degrees ), L5-S1 (25.8 degrees +/- 5.2 degrees and 20.8 degrees +/- 6 degrees ), and L1-S1 (53.4 degrees +/- 11.9 degrees and 57.3 degrees +/- 16.7 degrees ). Negative values reflect kyphosis, and positive values reflect lordosis. A significant difference between values before and during compression was obtained at L3-L4 and L5-S1. There was a significant decrease in disc height only at L4-L5 during compression. CONCLUSIONS: The axial force of 50% body weight in supine posture simulates the upright lumbar spine morphologically. No change in intervertebral angle occurred at L4-L5. However, disc height at L4-L5 decreased significantly during compression.

Non-NASA Center↗

Influence of Compression and Shear on the Strength of Composite Laminates with Z-Pinned Reinforcement

The influence of compression and shear loads on the strength of composite laminates with z-pins is evaluated parametrically using a 2D Finite Element Code (FLASH). Meshes were generated for three unique combinations of z-pin diameter and density. A laminated plate theory analysis was performed on several layups to determine the bi-axial stresses in the zero degree plies. These stresses, in turn, were used to determine the magnitude of the relative load steps prescribed in the FLASH analyses. Results indicated that increasing pin density was more detrimental to in-plane compression strength than increasing pin diameter. FLASH results for lamina with z-pins were consistent with the closed form results, and FLASH results without z-pins, if the initial fiber waviness due to z-pin insertion was added to the fiber waviness in the material to yield a total misalignment. Addition of 10% shear to the compression loading significantly reduced the lamina strength compared to pure compression loading. Addition of 50% shear to the compression indicated shear yielding rather than kink band formation as the likely failure mode. Two different stiffener reinforced skin configurations with z-pins, one quasi-isotropic and one orthotropic, were also analyzed. Six unique loading cases ranging from pure compression to compression plus 50% shear were analyzed assuming material fiber waviness misalignment angles of 0, 1, and 2 degrees. Compression strength decreased with increased shear loading for both configurations, with the quasi-isotropic configuration yielding lower strengths than the orthotropic configuration.

O'Brien, T. Kevin↗

Metronome Use for Coordination of Breaths and Cardiac Compressions Delivered by Minimally-Trained Caregivers During Two-Person CPR

Astronaut crew medical officers (CMO) aboard the International Space Station (ISS) receive 40 hours of medical training over 18 months before each mission, including two-person cardiopulmonary resuscitation (2CPR) as recommended by the American Heart Association (AHA). Recent studies have concluded that the use of metronomic tones improves the coordination of 2CPR by trained clinicians. 2CPR performance data for minimally-trained caregivers has been limited. The goal of this study was to determine whether use of a metronome by minimally-trained caregivers (CMO analogues) would improve 2CPR performance. 20 pairs of minimally-trained caregivers certified in 2CPR via AHA guidelines performed 2CPR for 4 minutes on an instrumented manikin using 3 interventions: 1) Standard 2CPR without a metronome [NONE], 2) Standard 2CPR plus a metronome for coordinating compression rate only [MET], 3) Standard 2CPR plus a metronome for coordinating both the compression rate and ventilation rate [BOTH]. Caregivers were evaluated for their ability to meet the AHA guideline of 32 breaths-240 compressions in 4 minutes. All (100%) caregivers using the BOTH intervention provided the required number of ventilation breaths as compared with the NONE caregivers (10%) and MET caregivers (0%). For compressions, 97.5% of the BOTH caregivers were not successful in meeting the AHA compression guideline; however, an average of 238 compressions of the desired 240 were completed. None of the caregivers were successful in meeting the compression guideline using the NONE and MET interventions. This study demonstrates that use of metronomic tones by minimally-trained caregivers for coordinating both compressions and breaths improves 2CPR performance. Meeting the breath guideline is important to minimize air entering the stomach, thus decreasing the likelihood of gastric aspiration. These results suggest that manifesting a metronome for the ISS may augment the performance of 2CPR on orbit and thus may increase the level of care.

Hurst, Victor, IV↗

Image compression system and method having optimized quantization tables

A digital image compression preprocessor for use in a discrete cosine transform-based digital image compression device is provided. The preprocessor includes a gathering mechanism for determining discrete cosine transform statistics from input digital image data. A computing mechanism is operatively coupled to the gathering mechanism to calculate a image distortion array and a rate of image compression array based upon the discrete cosine transform statistics for each possible quantization value. A dynamic programming mechanism is operatively coupled to the computing mechanism to optimize the rate of image compression array against the image distortion array such that a rate-distortion-optimal quantization table is derived. In addition, a discrete cosine transform-based digital image compression device and a discrete cosine transform-based digital image compression and decompression system are provided. Also, a method for generating a rate-distortion-optimal quantization table, using discrete cosine transform-based digital image compression, and operating a discrete cosine transform-based digital image compression and decompression system are provided.

Ratnakar, Viresh↗

Improving 3D Wavelet-Based Compression of Hyperspectral Images

Two methods of increasing the effectiveness of three-dimensional (3D) wavelet-based compression of hyperspectral images have been developed. (As used here, images signifies both images and digital data representing images.) The methods are oriented toward reducing or eliminating detrimental effects of a phenomenon, referred to as spectral ringing, that is described below. In 3D wavelet-based compression, an image is represented by a multiresolution wavelet decomposition consisting of several subbands obtained by applying wavelet transforms in the two spatial dimensions corresponding to the two spatial coordinate axes of the image plane, and by applying wavelet transforms in the spectral dimension. Spectral ringing is named after the more familiar spatial ringing (spurious spatial oscillations) that can be seen parallel to and near edges in ordinary images reconstructed from compressed data. These ringing phenomena are attributable to effects of quantization. In hyperspectral data, the individual spectral bands play the role of edges, causing spurious oscillations to occur in the spectral dimension. In the absence of such corrective measures as the present two methods, spectral ringing can manifest itself as systematic biases in some reconstructed spectral bands and can reduce the effectiveness of compression of spatially-low-pass subbands. One of the two methods is denoted mean subtraction. The basic idea of this method is to subtract mean values from spatial planes of spatially low-pass subbands prior to encoding, because (a) such spatial planes often have mean values that are far from zero and (b) zero-mean data are better suited for compression by methods that are effective for subbands of two-dimensional (2D) images. In this method, after the 3D wavelet decomposition is performed, mean values are computed for and subtracted from each spatial plane of each spatially-low-pass subband. The resulting data are converted to sign-magnitude form and compressed in a manner similar to that of a baseline hyperspectral- image-compression method. The mean values are encoded in the compressed bit stream and added back to the data at the appropriate decompression step. The overhead incurred by encoding the mean values only a few bits per spectral band is negligible with respect to the huge size of a typical hyperspectral data set. The other method is denoted modified decomposition. This method is so named because it involves a modified version of a commonly used multiresolution wavelet decomposition, known in the art as the 3D Mallat decomposition, in which (a) the first of multiple stages of a 3D wavelet transform is applied to the entire dataset and (b) subsequent stages are applied only to the horizontally-, vertically-, and spectrally-low-pass subband from the preceding stage. In the modified decomposition, in stages after the first, not only is the spatially-low-pass, spectrally-low-pass subband further decomposed, but also spatially-low-pass, spectrally-high-pass subbands are further decomposed spatially. Either method can be used alone to improve the quality of a reconstructed image (see figure). Alternatively, the two methods can be combined by first performing modified decomposition, then subtracting the mean values from spatial planes of spatially-low-pass subbands.

Klimesh, Matthew↗

Custom Gradient Compression Stockings May Prevent Orthostatic Intolerance in Astronauts After Space Flight

Orthostatic intolerance after space flight is still an issue for astronauts as no in-flight countermeasure has been 100% effective. NASA astronauts currently wear an inflatable anti-gravity suit (AGS) during re-entry, but this device is uncomfortable and loses effectiveness upon egress from the Shuttle. We recently determined that thigh-high, gradient compression stockings were comfortable and effective after space flight, though to a lesser degree than the AGS. We also recently showed that addition of splanchnic compression to this thigh-high compression stocking paradigm improved orthostatic tolerance to a level similar to the AGS, in a ground based model. Purpose: The purpose of this study was to evaluate a new, three-piece breast-high gradient compression garment as a countermeasure to post-space flight orthostatic intolerance. Methods: Eight U.S. astronauts have volunteered for this experiment and were individually fitted for a three-piece, breast-high compression garment to provide 55 mmHg compression at the ankle which decreased to approximately 20 mmHg at the top of the leg and provides ~15 mmHg over the abdomen. Orthostatic testing occurred 30 days pre-flight (w/o garment) and ~2 hours after flight (w/ garment) on landing day. Blood pressure (BP), Heart Rate (HR) and Stroke Volume (SV) were acquired for 2 minutes while the subject lay prone and then for 3.5 minutes after the subject stands up. To date, two astronauts have completed pre- and post-space flight testing. Data are mean SD. Results: BP [pre (prone to stand): 137+/-1.6 to 129+/-2.5; post: 130+/-2.4 to 122+/-1.6 mmHg] and SV [pre (prone to stand): 61+/-1.6 to 38+/-0.2; post: 58+/-6.4 to 37+/-6.0 ml] decreased with standing, but no differences were seen post-flight w/ compression garments compared to pre-flight w/o garments. HR [pre (prone to stand): 66+/-1.6 to 74+/-3.0, post: 67+/-5.6 to 78+/-6.8 bpm] increased with standing, but no differences were seen pre- to post-flight. Conclusion: After space flight, blood pressure and stroke volume are normally decreased and heart rate is usually elevated to compensate. In this small group of subjects, breast-high gradient compression stockings seem to have prevented these negative effects of spaceflight.

Stenger, Michael B.↗

Fast and Adaptive Lossless Onboard Hyperspectral Data Compression System

Modern hyperspectral imaging systems are able to acquire far more data than can be downlinked from a spacecraft. Onboard data compression helps to alleviate this problem, but requires a system capable of power efficiency and high throughput. Software solutions have limited throughput performance and are power-hungry. Dedicated hardware solutions can provide both high throughput and power efficiency, while taking the load off of the main processor. Thus a hardware compression system was developed. The implementation uses a field-programmable gate array (FPGA). The implementation is based on the fast lossless (FL) compression algorithm reported in Fast Lossless Compression of Multispectral-Image Data (NPO-42517), NASA Tech Briefs, Vol. 30, No. 8 (August 2006), page 26, which achieves excellent compression performance and has low complexity. This algorithm performs predictive compression using an adaptive filtering method, and uses adaptive Golomb coding. The implementation also packetizes the coded data. The FL algorithm is well suited for implementation in hardware. In the FPGA implementation, one sample is compressed every clock cycle, which makes for a fast and practical realtime solution for space applications. Benefits of this implementation are: 1) The underlying algorithm achieves a combination of low complexity and compression effectiveness that exceeds that of techniques currently in use. 2) The algorithm requires no training data or other specific information about the nature of the spectral bands for a fixed instrument dynamic range. 3) Hardware acceleration provides a throughput improvement of 10 to 100 times vs. the software implementation. A prototype of the compressor is available in software, but it runs at a speed that does not meet spacecraft requirements. The hardware implementation targets the Xilinx Virtex IV FPGAs, and makes the use of this compressor practical for Earth satellites as well as beyond-Earth missions with hyperspectral instruments.

Aranki, Nazeeh I.↗

A New Compression Method for FITS Tables

As the size and number of FITS binary tables generated by astronomical observatories increases, so does the need for a more efficient compression method to reduce the amount disk space and network bandwidth required to archive and down1oad the data tables. We have developed a new compression method for FITS binary tables that is modeled after the FITS tiled-image compression compression convention that has been in use for the past decade. Tests of this new method on a sample of FITS binary tables from a variety of current missions show that on average this new compression technique saves about 50% more disk space than when simply compressing the whole FITS file with gzip. Other advantages of this method are (1) the compressed FITS table is itself a valid FITS table, (2) the FITS headers remain uncompressed, thus allowing rapid read and write access to the keyword values, and (3) in the common case where the FITS file contains multiple tables, each table is compressed separately and may be accessed without having to uncompress the whole file.

Pence, William↗

Determining Compression Characteristics of Honeycomb Material - 19674

The objective is to determine compression test characteristics of stainless steel honeycomb material to be able to represent honeycomb structures accurately in analytical models used to simulate hypothetical accident scenarios of shipping packages. Honeycomb is a material used primary in the aerospace industry due to its high strength-to-weight ratio. Because it doesn't have a shelf life and can withstand high heats, it is an excellent candidate for a structural material in package designs. Honeycomb has orthotropic material properties. The honeycomb currently being evaluated is made of metal ribbons spot welded together to form hexagon pattern between metal plates. The hexagon ribbons are brazed to the metal plates. The complexity of the honeycomb significantly lengthens the simulation time, which is further complicated when the brazing, welding, and imperfections of the material is considered. This means to affectively represent the honeycomb in simulation programs, such as Abaqus, the material needs to be approximated as a uniform orthotropic material. This requires material properties in each of the three directions, T, W, and L. The T direction is defined as the 'strong' direction, perpendicular to the honeycomb sheet. The L direction is parallel the ribbon and the W direction is perpendicular to the ribbon. Honeycomb has 3 different compression stages. Stage 1 is the initial compression where the honeycomb maintains its structural integrity and does not permanently deform from loads from normal operations. Stage 2 is when the initial buckling causes irreversible damage to the honeycomb. Stage 2 is the one we are most interested in because it absorbs the most energy from a hypothetical accident scenario. During Stage 2, the honeycomb fails layer-by-layer, indicating that the more layers, the longer material crushing is sustained, and thus the more energy that is absorbed. Stage 3 is final compression, similar to compressing solid metal. During stage 3 the stress increases with a diminishing rate of elongation and the honeycomb is completely failed where the plates between the honeycomb core sandwich the crushed honeycomb ribbon. The stress-strain graph bellow shows compression test of 3 different samples in the T direction. The red lines divide the three stages. The far left is Stage 1, the middle stage 2, and the right stage 3. The data collected is displayed in the table below. The expressions in the table represent the regression formula that represents each stage of the compression on a stress-strain graph. Stage 1 is assumed to intersect with the origin; data for stage 3 for the crush in W and L directions where unable to be gathered due to the nature of failure for those directions. Stage 1 and 2 are linear regressions while stage 3 is a degree 2, to best match the curve. Linear regression is used on stage 2 because ultimately that represents the energy absorbed. With more data, a degree n x 2 regression would be a more appropriate regression for stage 2, where n is the number of layers. The table below is the approximation of each scenario and stage. Each stage starts and ends at the intersection of the next stage. There where two main objectives for these tests: firstly to understand how honeycomb performs under extreme compression, and secondly to be able to numerically represent the honeycomb structure. Both of these where completed to varying degrees. It is important understood how the honeycomb fails. If it is crushed in the T direction it retains its integrity even after being crushed, however when crushed in the L and W crush direction, if it fails, it disintegrates, and loses all integrity. Also, the more layers the more time the material spends in stage 2. Failure is started by buckling, thus if there is any imperfection, the stress will not spike but transition straight into stage 2. The data collected and aggregated can be used for initial simulation of honeycomb used in packages. The initial testing has set the ground work for more data to be collected in order to verify results and to allow more confidence in the simulation results. This is only the initial data collected. The next steps is to continue to collect more data to verify results and to test more variants of honeycomb, with different brazing, layers, and shape.

36 MATERIALS SCIENCE↗

Strategies for on-chip digital data compression for X-ray pixel detectors

Here, the continued desire for X-ray pixel detectors with higher frame rates will stress the ability of application-specific integrated circuit (ASIC) designers to provide sufficient off-chip bandwidth to reach continuous frame rates in the 1 MHz regime. To move from the current 10 kHz to the 1 MHz frame rate regime, ASIC designers will continue to pack as many power-hungry high-speed transceivers at the periphery of the ASIC as possible. In this paper, however, we present new strategies to make the most efficient use of the off-chip bandwidth by utilizing data compression schemes for X-ray photon-counting and charge-integrating pixel detectors. In particular, we describe a novel in-pixel compression scheme that converts from analog to digital converter units to encoded photon counts near the photon Poisson noise level and achieves a compression ratio of >1.5x independent of the dataset. In addition, we describe a simple yet efficient zero-suppression compression scheme called "zeromask" (ZM) located at the ASIC's edge before streaming data off the ASIC chip. ZM achieves average compression ratios of >4x, >7x, and >8x for high-energy X-ray diffraction, ptychography, and X-ray photon correlation spectroscopy datasets, respectively. We present the conceptual designs, register-transfer level block diagrams, and the physical ASIC implementation of these compression schemes in 65 nm CMOS. When combined, these two digital compression schemes could increase the effective off-chip bandwidth by a factor of 6-12x.

47 OTHER INSTRUMENTATION↗

FedCSpc: A Cross-Silo Federated Learning System With Error-Bounded Lossy Parameter Compression

Cross-Silo federated learning is widely used for scaling deep neural network (DNN) training over data silos from different locations worldwide while guaranteeing data privacy. Communication has been identified as the main bottleneck when training large-scale models due to large-volume model parameters and gradient transmission across public networks with limited bandwidth. Most previous works focus on gradient compression, while limited work tries to compress parameters that can not be ignored and extremely affect communication performance during the training. Here, to bridge this gap, we propose FedCSpc: an efficient cross-silo federated learning system with an XAI-driven adaptive parameter compression strategy for large-scale model training. Our work substantially differs from existing gradient compression techniques due to the distinct data features of gradient and parameter. The key contributions of this paper are fourfold. (1) Our designed FedCSpc proposes to compress the parameter during the training using the state-of-the-art error-bounded lossy compressor – SZ3. (2) We develop an adaptive compression error bound adjustment algorithm to guarantee the model accuracy effectively. (3) We exploit an efficient approach to utilize the idle CPU resources of clients to compress the parameters. (4) We perform a comprehensive evaluation with a wide range of models and benchmarks on a GPU cluster with 65 GPUs. Results show that FedCSpc can achieve the same model accuracy as FedAvg while reducing the data volume of parameters and gradients in communication by up to 7.39× and 288×, respectively. With 32 clients on a 4 Gb size model, FedCSpc significantly outperforms FedAvg in wall-clock time in the emulated WAN environment (at the bandwidth of 1 Gbps or lower without loss of generality).

SZ3↗

Transported PDF Modeling of Compressible Turbulent Reactive Flows by using the Eulerian Monte Carlo Fields Method

Although the transported probability density function (PDF) method has been developed for decades, its application has been mainly focused on the low-Mach number flow problems. This work extends the transported PDF method to compressible flow problems. The Eulerian Monte Carlo fields (EMCF) solution method is employed to solve the transported PDF equation for compressible flow problems. A pseudo stagnation enthalpy is introduced and its stochastic partial differential equation is derived to ensure total energy conservation numerically. A new mixing model called interaction by partial exchange with mean (IPEM) is introduced to expand the available choices of mixing models for the EMCF method. The consistency of the EMCF method is examined for solving the transported PDF equation. Numerical implementation details are discussed, such as the density coupling between the compressible flow solver and the EMCF solver, discretization schemes for the mixing terms and the stochastic terms. The implemented compressible flow solver coupled with the EMCF solver is verified and validated in a series of test cases with increasing level of complexity, ranging from a statistically one-dimensional turbulent mixing layer to a self-excited resonance model rocket combustor. It is observed that in general with the increase of compressibility, there is an increase in the sensitivity of the modeling results to the different models and algorithms. This makes it necessary to develop a thorough understanding of the model sensitivity in order to develop a robust and accurate simulation solver for highly compressible turbulent reactive flows. The thermo-acoustic instability inside the model rocket combustor case is captured reasonably, which demonstrates the overall capability of the developed compressible turbulent combustion solver based on the transported PDF method.

42 ENGINEERING↗

Sound speed measurements in shock-compressed cemented Tungsten carbides at pressures up to 100 GPa

To gain insights into thermodynamic states attained during shock compression of cemented tungsten carbide with 3.7 wt.% cobalt binder, we present results of longitudinal sound (release wave) speed measurements and their analysis at peak stresses up to 100 GPa (volumetric compression ratio ~ 15%). The sound speeds are determined using front-surface impact and release-wave overtake plate impact experimental configurations using laser interferometry. The measured sound speed data along with estimates for bulk sound speeds obtained using the fourth-order Birch-Murnaghan EoS and thermodynamics are used to determine the longitudinal moduli and shear moduli of shocked tungsten carbide at the various peak compression states attained in the experiments. Here, the longitudinal sound speeds were found to increase linearly with volume compression ratio from 6.97 ± 0.010 km/s at ambient conditions to 8.26 ± 0.156 km/s at a volume compression ratio of ~ 15%. The corresponding longitudinal elastic moduli also increase nearly linearly with the volume compression ratio but remain consistently lower than their theoretical predictions based on continuum models with no damage. Also, the sensitivity of shear moduli to pressure, as predicted by the Steinberg-Guinan model, is reduced substantially and the shear moduli of cemented WC with 3.7 wt.% Co remains nearly constant at ~ 310 GPa at the various peak compression stress states investigated in the present study.

36 MATERIALS SCIENCE↗

A 28 nm multiply-accumulate ASIC architecture for on-chip data compression in MHz frame rate X-ray and electron pixel detectors

Modern X-ray detector systems urgently require compact, efficient, and fast data compression schemes to handle the transmission of big data from pixel arrays, enabling frame rates in the MHz regime. Here, in this work, a data compression ASIC that implements a streaming fixed-length lossy compression scheme is introduced and analyzed, proving the feasibility and benefits of on-chip compression. The compression scheme utilizes a vector matrix product logic, which performs a number of floating-point multiplications, additions, and accumulations. The logic is verified, synthesized, and shown to fit in the area resource available for the X-ray detector under study, which comprises 192 × 168 pixels each of 12-bit width, and having a total area of 20 mm× 20 mm, about 2 mm× 20 mm of which are available for the digital logic. Several system architectures, precisions, and compression ratios ranging from 100 to 250 were analyzed to pave the way for on-chip fixed-length compression (e.g., principal component analysis, singular value decomposition) and data reduction (e.g., azimuthal integration) for X-ray and electron detectors.

Data compression↗

zPerf: A Statistical Gray-Box Approach to Performance Modeling and Extrapolation for Scientific Lossy Compression

With the scaling up of simulation-based scientific discovery on high-performance computing systems, the disparity between compute and I/O has increased, forcing domain scientists to save only a small amount of simulation data to persistent storage. This can result in the loss of essential physics fields that are needed for data analysis. While error-bounded lossy compression has made tremendous progress in bridging the gap between compute and I/O, the lack of understanding of compression performance remains a key hurdle to its wide adoption. Here, in this work, we present zPerf, a statistical gray-box performance modeling approach for scientific lossy compression. Our contributions are threefold: 1) We develop zPerf to estimate the performance of lossy compression techniques, based on in-depth understanding and statistical modeling for data features and core compression metrics; 2) We demonstrate the in-detailed implementation of zPerf using two case studies, where we derive the performance modeling for SZ and ZFP, two leading lossy compressors; 3) We evaluate the effectiveness of zPerf on real-world datasets across various domains. Based on the evaluation, we demonstrate the efficacy of the zPerf performance model; 4) We further discuss three case studies where zPerf is applied to extrapolate the compression ratio of SZ and ZFP with alternative encoding schemes as well as ZFP with an alternative transform scheme. Through the case studies, we demonstrate the potential of zPerf for exploring the design space of lossy compression, which has hardly been studied in the literature.

97 MATHEMATICS AND COMPUTING↗

Optimizing Error-Bounded Lossy Compression for Scientific Data With Diverse Constraints

Vast volumes of data are produced by today's scientific simulations and advanced instruments. These data cannot be stored and transferred efficiently because of limited I/O bandwidth, network speed, and storage capacity. Error-bounded lossy compression can be an effective method for addressing these issues: not only can it significantly reduce data size, but it can also control the data distortion based on user-defined error bounds. In practice, many scientific applications have specific requirements or constraints for lossy compression, in order to guarantee that the reconstructed data are valid for post hoc analysis. For example, some datasets contain irrelevant data that should be isolated in particular and users often have intuition regarding value ranges, geospatial regions, and other data subsets that are crucial for subsequent analysis. Existing state-of-the-art error-bounded lossy compressors, however, do not consider these constraints during compression, resulting in inferior compression ratios with respect to user's post hoc analysis, due to the fact that the data itself provides little or no value for post hoc analysis. In this work we address this issue by proposing an optimized framework that can preserve diverse constraints during the error-bounded lossy compression, e.g., cleaning the irrelevant data, efficiently preserving different precision for multiple value intervals, and allowing users to set diverse precision over both regular and irregular regions. We perform our evaluation on a supercomputer with up to 2,100 cores. Experiments with six real-world applications show that our proposed diverse constraints based error-bounded lossy compressor can obtain a higher visual quality or data fidelity on reconstructed data with the same or even higher compression ratios compared with the traditional state-of-the-art compressor SZ. Furthermore, our experiments also demonstrate very good scalability in compression performance compared with the I/O throughput of the parallel file system.

97 MATHEMATICS AND COMPUTING↗

Stability Analysis of Inline ZFP Compression for Floating-Point Data in Iterative Methods

Currently, the dominating constraint in many high performance computing applications is data capacity and bandwidth, in both internode communications and even moreso in intranode data motion. A new approach to address this limitation is to make use of data compression in the form of a compressed data array. Storing data in a compressed data array and converting to standard IEEE-754 types as needed during a computation can reduce the pressure on bandwidth and storage. However, repeated conversions (lossy compression and decompression) introduce additional approximation errors, which need to be shown to not significantly affect the simulation results. Here, we extend recent work that analyzed the error of a single use of compression and decompression of the ZFP compressed data array representation to the case of time-stepping and iterative schemes, where an advancement operator is repeatedly applied in addition to the conversions. We show that the accumulated error for iterative methods involving fixed-point and time evolving iterations is bounded under standard constraints. An upper bound is established on the number of additional iterations required for the convergence of stationary fixed-point iterations. An additional analysis of traditional forward and backward error of stationary iterative methods using ZFP compressed arrays is also presented. The results of several 1D, 2D, and 3D test problems are provided to demonstrate the correctness of the theoretical bounds.

97 MATHEMATICS AND COMPUTING↗