Search NASASearch

NASA NTRS · 19940009906

Estimating the Size of Huffman Code Preambles

Abstract

Data compression via block-adaptive Huffman coding is considered. The compressor consecutively processes blocks of N data symbols, estimates source statistics by computing the relative frequencies of each source symbol in the block, and then synthesizes a Huffman code based on these estimates. In order to let the decompressor know which Huffman code is being used, the compressor must begin the transmission of each compressed block with a short preamble or header file. This file is an encoding of the list n = (n 1 , n 2 ....,n m ), where n i is the length of the Hufffman codeword associated with the ith source symbol. A simple method of doing this encoding is to individually encode each n i into a fixed-length binary word of length log 2 l, where l is an a priori upper bound on the codeword length. This method produces a maximum preamble length of mlog 2 l bits. The object is to show that, in most cases, no substantially shorter header of any kind is possible.

Keep this discovery

BibTeXRIS

R J McEliece, T H Palmatier. 1993-08-15. Estimating the Size of Huffman Code Preambles. https://ntrs.nasa.gov/citations/19940009906

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related reports

PERSIANN-Unet: A Global Deep Learning Framework for Near-Real-Time Precipitation Estimation Using Infrared Data

Access to high-quality, high-resolution, near-real-time precipitation data is essential for hydrological and meteorological research and disaster mitigation. Traditional tools such as rain gauges and radar networks, though effective, have limitations, including sparse coverage in remote areas and high operational costs. Satellite data, with its global coverage and high spatial and temporal resolutions, mitigates limitations in coverage. Satellite precipitation products like Hydro Estimator (HE), Integrated Multi-satellitE Retrievals for Global Precipitation Measurement (IMERG), and Precipitation Estimation from Remotely Sensed Information using Artificial Neural Networks (PERSIANN) utilize both geosynchronous thermal infrared (IR) and passive microwave (PMW) data in their operation. PMW sensors offer detailed atmospheric profiles but suffer from higher latency, whereas IR sensors provide lower latency but only capture cloud-top information. Despite this constraint, IR data remains attractive for low-latency precipitation estimation. Recent advances in deep learning, particularly convolutional neural networks (CNNs), have further improved satellite precipitation retrievals. This study introduces PERSIANN-Unet (PUnet or PERSIANN V3), a quasi-global algorithm covering 60°N–60°S that combines IR data, monthly climatology, and the UNet architecture to produce half-hourly precipitation estimates at 0.04° resolution. The product is evaluated against HE, IMERG, and PDIR-Now for 2022–2023. Results show that PUnet closely matches its training target, IMERG V07 Final, at the global scale, and performance is further evaluated against Stage IV as a reference over CONUS. Training PUnet on IMERG (2016–2021) leverages a high-quality, integrated PMW IR-gauge precipitation product while developing an IR-based framework not reliant on PMW availability. By operating on a single global image, PUnet avoids tile partitioning and blending steps, reducing edge discontinuities, and produces more spatially consistent precipitation fields across hemispheres.

Phu Nguyen

Numerical-heating effects in atmospheric pressure streamer discharges simulated with a PIC code

Artificial heating in plasma simulations is a well-known phenomenon which occurs when, among other things, the Debye length is poorly resolved by the simulation mesh. Here, in this work, the degree to which numerical-heating occurs during a simulation of a nanosecond atmospheric pressure streamer discharge is examined. The streamer is simulated using a two-dimensional finite-element, particle-in-cell code Empire, which uses direct simulation Monte Carlo for binary particle interactions. Initially, an estimate of the numerical-heating rate applied to Empire is performed using a simple plasma model. Second, a positive atmospheric pressure streamer discharge simulation is performed to study the effects of numerical heating on plasma density, electron temperature, and streamer velocity. The nominal Debye length is approximately 1 μm and the amount of numerical heating introduced in the simulation is varied by using mesh sizes ranging from 2 μm to 20 μm. A measurable numerical heating quantity is proposed that can be used to estimate the appropriate element size and quantify the numerical-heating that can be expected over the simulation time for an atmospheric pressure streamer. In conclusion while Δx/λ D violations can be an issue it is not likely to be an issue with streamer discharges that are temporally short and occur in environments where collision frequencies are high. This result validates the rationale of grid size choices for a large amount of previously published works where Δx/λ D violation was not clearly addressed. Primary finding of this work is that numerical heating is of minor concern for plasma simulations where electron–neutral collisions are numerous such that multiple collisions can occur within a single plasma period.

Nikic, Dejan [University of New Mexico, Albuquerqu

Incremental Learning for Passive Microwave Precipitation Retrievals using Advanced Technology Microwave Sounder

Spaceborne passive microwave (PMW) radiometry is central to global precipitation monitoring, yet retrieval uncertainties remain substantial, particularly for cross-track sounders whose variable footprints and channel configurations are optimized for atmospheric temperature and moisture profiling rather than precipitation. Consequently, existing operational products often exhibit angular-dependent biases, limited effective swath utilization, unrealistic rainfall probability distributions, and systematic misclassification of precipitation phase. These limitations are further compounded by the scarcity of globally accurate and representative precipitation observations, as training data from the Dual-frequency Precipitation Radar (DPR) and the Cloud Profiling Radar (CPR) are spatially sparse, lack uniform global coverage, and exhibit heterogeneous error characteristics across precipitation regimes. To address these challenges, this study presents a supervised retrieval algorithm that incrementally trains an ensemble of extreme gradient-boosted decision trees by augmenting base learners with pre-training on reanalysis data and post-training on coincident DPR and CPR observations matched with the Advanced Technology Microwave Sounder (ATMS). By transferring prior information from reanalysis to posterior constraints from radar observations and adopting a sequential detection–estimation strategy for precipitation phase and rate retrieval, the proposed approach yields retrievals across the full ATMS swath that are largely free from persistent deficiencies in current Global Precipitation Measurement (GPM) passive microwave operational products. In particular, the method resolves bimodal artifacts in rainfall retrievals and mitigates systematic high-latitude snowfall biases, including overestimation across the Arctic and underestimation across the Antarctic. Validation against independent Multi-Radar Multi-Sensor (MRMS) data over the Contiguous United States (CONUS) further demonstrates improved performance in precipitation phase detection and rate estimation relative to both reanalysis and current GPM PMW products.

Mahyar Garshasbi