WO2008021304A2 - Data encoder - Google Patents

Data encoder Download PDF

Info

Publication number
WO2008021304A2
WO2008021304A2 PCT/US2007/017893 US2007017893W WO2008021304A2 WO 2008021304 A2 WO2008021304 A2 WO 2008021304A2 US 2007017893 W US2007017893 W US 2007017893W WO 2008021304 A2 WO2008021304 A2 WO 2008021304A2
Authority
WO
WIPO (PCT)
Prior art keywords
word
circuit
data stream
bit
zero
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Ceased
Application number
PCT/US2007/017893
Other languages
French (fr)
Other versions
WO2008021304A3 (en
Inventor
James L. Fulcomer
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Raytheon Co
Original Assignee
Raytheon Co
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Raytheon Co filed Critical Raytheon Co
Priority to EP07811287.7A priority Critical patent/EP2055007B1/en
Publication of WO2008021304A2 publication Critical patent/WO2008021304A2/en
Publication of WO2008021304A3 publication Critical patent/WO2008021304A3/en
Priority to IL197045A priority patent/IL197045A/en
Anticipated expiration legal-status Critical
Ceased legal-status Critical Current

Links

Classifications

    • HELECTRICITY
    • H03ELECTRONIC CIRCUITRY
    • H03MCODING; DECODING; CODE CONVERSION IN GENERAL
    • H03M7/00Conversion of a code where information is represented by a given sequence or number of digits to a code where the same, similar or subset of information is represented by a different sequence or number of digits
    • H03M7/30Compression; Expansion; Suppression of unnecessary data, e.g. redundancy reduction
    • H03M7/40Conversion to or from variable length codes, e.g. Shannon-Fano code, Huffman code, Morse code

Definitions

  • the present invention relates to signal processing systems. More specifically, the present invention relates to data compression encoders.
  • Data compression is used in a variety of applications to encode data using fewer bits than the original representation in order to reduce the consumption of resources such as storage space or transmission bandwidth. Lossless data compression accomplishes this without any loss of information; that is, the original data can be reconstructed exactly from the encoded data.
  • Lossless data compression algorithms typically include two sections: a preprocessor, which transforms the input data using a statistical model into samples that can be more efficiently compressed (so that certain symbols occur more frequently than others), and an encoder, which encodes the transformed data using a scheme such that more probable symbols produce shorter output than improbable symbols. Entropy encoders encode symbols such that the code length is proportional to the probability of the symbol. More common symbols therefore use the shortest codes.
  • a Rice encoder divides an input word into two variable-length sections. One section is encoded using fundamental sequence encoding, and the other section is left alone (remains binary).
  • a Rice encoder is usually implemented in software running on a computer system. This implies high power consumption, weight, size, and recurring cost. Certain applications, particularly space or airborne applications, have size, weight, and power constraints that would prohibit the use of a large computer system. In addition, some applications require that the encoded data is output at a rate matching that of the incoming data. These applications require an encoder that can operate at faster processing speeds than can be achieved with conventional software implementations.
  • a hardware approach could potentially offer faster processing speeds as well as smaller size, weight, and power consumption; however, conventional encoder architectures are either too large to realize in current digital technologies, or are too slow (i.e., output rates are slower than input rates). Hence, a need exists in the art for an improved encoder offering smaller size, weight, and power consumption, as well as faster processing speeds than conventional implementations.
  • the novel encoder includes a first circuit for generating a fundamental sequence coded data stream from an incoming input data stream, a second circuit for generating a k-split data stream from the incoming data stream, and a third circuit for combining the fundamental sequence coded data stream and k-split data stream to form a final encoded output.
  • the first circuit includes a circuit for converting the incoming input data stream into a novel intermediate format comprising a set bit word and a zero word count, and a zero-word expander for converting the intermediate format to the fundamental sequence coded data stream.
  • the encoder may also include a register adapted to store the intermediate format to provide rate buffering to allow for full speed input and output rates.
  • Fig. 1 is a simplified block diagram of a lossless data compression system designed in accordance with an illustrative embodiment of the present invention.
  • Fig. 2 is a diagram showing an illustrative coded data set format.
  • Fig. 3 is a simplified block diagram of an illustrative embodiment of an entropy encoder designed in accordance with the present teachings.
  • Fig. 1 is a simplified block diagram of a lossless data compression system 10 designed in accordance with an illustrative embodiment of the present invention.
  • the system 10 includes a preprocessor 12 and a novel entropy encoder 14.
  • the preprocessor 12 receives the input data x and transforms it into preprocessed data ⁇ suitable for the entropy encoding (typically reducing the entropy of the data stream).
  • the entropy encoder 14 then converts the preprocessed data ⁇ into an encoded bit sequence y.
  • the input data x is processed in blocks of J /i-bit words:
  • x,- is an ⁇ -bit word and n is a constant.
  • the preprocessor transforms the input data x into blocks of preprocessed samples ⁇ :
  • ⁇ / is an ⁇ -bit integer between 0 and 2"-l.
  • the preprocessed data stream is weighted heavily towards zero.
  • One simple method for achieving this weighting is to make a prediction of the next value in the data stream based on the current value, and then transmit the difference between predicted and actual.
  • the preprocessed data stream ⁇ would be ⁇ 100, -1, -1, +1 ⁇ . Note that while the first value of this stream is a large number, the subsequent values are much smaller.
  • the preprocessor may also output a code ID and/or reference data that can be used in a decoding stage to reverse the preprocessor transform function and recover the original input data.
  • the system 10 uses a Rice encoding algorithm.
  • the Rice algorithm splits each preprocessed sample ⁇ ,- into two sections. The k least significant bits are split off from each sample and the remaining bits (the n-k most significant bits) are encoded using a fundamental sequence (FS) codeword.
  • the variable k can vary for different blocks of J words (all samples within a block are encoded using the same k split).
  • the preprocessor selects a value for k that optimizes compression for that block, and outputs a code ID field that indicates the value of A:.
  • Fundamental sequence encoding uses variable-length codewords, where an integer value m is represented by m zeros followed by a one.
  • the following table shows the FS codeword for various values of the preprocessed sample ⁇ ,-:
  • FS codewords from each of the J preprocessed words are concatenated to form a single sequence, which is output along with the removed k-split bits.
  • the final encoded output y includes a coded data set (CDS) for each block of J input words.
  • Fig. 2 is a diagram showing an illustrative coded data set format.
  • Each CDS includes a code ID field that indicates the value of k, an optional reference field that includes information from the preprocessor used to reverse the preprocessor transformation, an FS field that includes the concatenated FS codewords from all J words, and a k-split field including all of the k-split least significant bits removed from the J words.
  • the entropy encoder is usually implemented in software. Certain applications would benefit from a hardware solution that can offer smaller size, weight and power consumption, as well as faster processing speeds.
  • the present invention provides a novel entropy encoder architecture suitable for hardware implementations.
  • the novel algorithm disassembles the uncoded input data stream into two separate streams (a fundamental sequence stream and a k-split stream) that are each partially pre-coded and partially assembled, then later processed and recombined into the final output.
  • Fig. 3 is a simplified block diagram of an illustrative embodiment of an entropy encoder 14 designed in accordance with the present teachings.
  • the novel encoder 14 includes a code ID/reference FIFO (first in, first out register) 20 for storing the code ID and optional reference fields, a k-split generator 22 for computing, buffering and assembling the k-split output stream, a fundamental sequence generator 24 for computing and assembling the FS output stream, and an output assembly manager 26 for assembling the code ID and reference fields, the FS output stream, and the k-split output stream into a final coded data set.
  • the FS generator 24 and k-split generator 22 operate in parallel to maintain full-speed input and output.
  • the k-split generator 22 includes a k-split packer 50 and a k-split FIFO 52.
  • the k-split packer 50 accepts the incoming stream of uncoded data ⁇ and packs the k split bits of each data word ⁇ , together into a sequence of M-bit words.
  • the k-split packer 22 includes a multi-bit shifter/masker 42 and a k-split accumulator 44.
  • the multi-bit shifter/masker 42 is a single cycle or pipelinable shifter capable of different shift values every clock cycle.
  • the shifter/masker 42 receives each it-bit uncoded input data word ⁇ , and the value of k, and masks out the n—k most significant bits, keeping only the k least significant bits.
  • the Ar-split bits are shifted and accumulated in the k-split accumulator 44 to form an M-bit word comprising the k-split bits from one or more data words packed together end-to-end.
  • the M-bit word is sent to the k-split FIFO 52, and the next k-split bits are accumulated in the next M-bit word.
  • the FIFO 52 is large enough to store the prepacked k-split data for at least one full CDS, thus allowing the final packing and output stage to stream data out at full speed.
  • M is 16 bits; however, other word lengths can be used without departing from the scope of the present teachings.
  • the FS generator 24 uses a novel FS encoding technique that includes transforming the incoming stream of uncoded data into an intermediate format.
  • the novel fundamental sequence intermediate format (FSIF) includes two fields: a pre- packed set bit word and a zero word count.
  • the set bit word is an L-bit word with one or more bits set to '1', each set bit representing a full or partial uncoded data word.
  • the zero word count represents the number of all-zero words to be inserted between the pre-packed set bit words.
  • This transformation prepares the incoming data for easy conversion to full fundamental sequence coding while allowing the input stream to be accepted at full speed.
  • the intermediate format data can be stored in a shallow FIFO to absorb any short term bursts of inefficient compression. The intermediate format data is then converted to full FS code and output at full speed.
  • the FS generator 24 includes a FSIF packer 30, a fundamental sequence FIFO 32, and a zero-word expander 34.
  • the FSIF packer 30 accepts the incoming stream of uncoded data ⁇ and transforms it into the intermediate format comprised of a pre-packed set bit word and a zero word count. This intermediate format is stored in the FIFO 32.
  • the zero-word expander 34 pulls the intermediate format data from the FIFO 32 and converts it to a fundamental sequence coded output by expanding the zero word count into L-bit all-zero words and merging them with the pre-packed set bit words.
  • the FSEF packer 30 includes a multi-bit shifter/masker 40, a set bit calculator 42, a set bit accumulator 44, and an FSIF distributor 46.
  • the multi-bit shifter/masker 40 receives each uncoded input data word ⁇ ,- and the value of k, and masks out the k least significant bits. The remaining n—k bit word ⁇ 't is sent to the set bit calculator 42.
  • the set bit calculator 42 includes an adder adapted to add the truncated data words ⁇ 'j together in a particular way and output the resulting sums using a special "one- hot" format, and logic for generating a zero word count.
  • a one-hot format word has one and only one bit set to '1'; all other bits are set to zero.
  • a value of/ encoded in one-hot would be all zeros except for a '1 ' in the b / position. Thus. a value of '2' would be encoded as ⁇ 000...000100 ⁇ .
  • the first sum Si output by the set bit calculator 42 is formed by adding the first truncated word ⁇ 'i to zero.
  • the result is output by the set bit calculator 42 in one-hot format. All subsequent sums s,- are formed by adding the incoming word ⁇ ',- to the previous sum s,--y plus one.
  • Each sum s,- is output in one-hot format to the set bit accumulator 44.
  • the set bit calculator 42 is also adapted to output a zero word count in addition to an L-bit one- hot encoded word, hi an illustrative embodiment, L is set to 16 bits (other word lengths can be used without departing from the scope of the present teachings).
  • L is set to 16 bits (other word lengths can be used without departing from the scope of the present teachings).
  • the calculator 42 tries to encode a number greater than L— 1 in one-hot format, the adder overflows.
  • the calculator therefore includes logic for determining when the adder overflows and by how much, i.e., by how many sets of L zeros. This zero word count is output to the FSIF distributor 46.
  • the set bit accumulator 44 accumulates the sums output from the set bit calculator 42 to form a pre-packed set bit word with one or more bits set to T, each set bit representing a full or partial uncoded data word.
  • the accumulator 44 starts with an L-bit word of all zeros.
  • Each word output by the set bit calculator 42 includes one and only one bit set to ' 1 '.
  • the accumulator 44 sets one of its bits (in the same position as the T bit in the word output from the calculator 42) to 1 V. This process is repeated for each word the accumulator 44 receives, until the calculator 42 overflows.
  • the accumulator 44 thus keeps track of all the ' 1 ' bits from one or more words.
  • the L-bit word in the accumulator 44 is output to the FSIF distributor 46.
  • the accumulator 44 then resets to zero and begins to accumulate the next L-bit word.
  • the FSIF distributor 46 is a timing and control function used to parse the set bit accumulator and set bit calculator outputs into the intermediate format.
  • the distributor 46 is adapted to receive the zero word count from the set bit calculator 42. When the zero word count indicates that the set bit calculator 42 has overflowed, the distributor 46 pulls the L-bit word from the accumulator 44 and outputs the L-bit word and the zero word count to the FIFO 32.
  • the accumulator 44 receives the sum si and accumulates it with its previous state (all zeros), resulting in ⁇ 0000000000000100 ⁇ .
  • the set bit calculator 42 receives a ' 1 ' and outputs the following sum encoded in one-hot format to the accumulator 44:
  • the accumulator 44 accumulates the new sum s 2 with its previous state, resulting in ⁇ 0000000000010100 ⁇ .
  • the set bit calculator 42 receives a '3' and outputs the following sum to the accumulator 44:
  • the set bit calculator logic therefore outputs a zero word count of '2'
  • the distributor 46 receives the zero word count indicating that the set bit calculator 42 has overflowed, it pulls the 16-bit word that was in the accumulator 44 (before the overflow). In this example, it pulls the word ⁇ 0000000100010100 ⁇ from the accumulator 44 and outputs the word and the zero word count to the FIFO 32.
  • the zero word count of '2' indicates that the zero-word expander 34 should output the 16-bit word from the accumulator, followed by one 16-bit word of all zeros.
  • the set bit calculator 42 receives a '2' and outputs the following sum to the accumulator 44:
  • the accumulator 44 accumulates the new sum S 5 with its previous state, resulting in ⁇ 0000100100000000 ⁇ . This process continues until all J words in the incoming data block ⁇ are processed.
  • the intermediate format data output from the distributor 46 is stored in the FIFO 32.
  • the zero-word expander 34 pulls the intermediate format data (the pre-packed set bit word and the zero word count) from the FIFO 32 and converts it to a full, fundamental sequence coded output.
  • the zero- word expander 34 includes a first circuit adapted to convert the zero word count to a number of L-bit words of all zeros and a second circuit adapted to output the set bit word followed by the all-zero words (if any).
  • a zero word count of ' 1 ' corresponds with no all-zero words, so the zero-word expander 34 just outputs the pre-packed set bit word and waits to receive the next set bit word and zero word count.
  • a zero word count of '2' corresponds with one all-zero word, so the zero-word expander 34 outputs the pre-packed set bit word followed by one L-bit word of all zeros.
  • the sequence of words output from the zero-word expander 34 forms the concatenated FS coded output.
  • the FIFO 32 provides internal rate buffering to allow full-speed input and output rates.
  • the FSIF packer 30 processes an input word ⁇ ',- having a large value (i.e., resulting in a zero word count greater than 1)
  • the output from the zero- word expander 34 will take more than one clock cycle: one cycle for outputting the pre-packed set bit word and an additional clock cycle for each all-zero word.
  • the FSIF packer 30 processes input words ⁇ ', having small values (i.e., much smaller than L-I)
  • several input words are accumulated into a single L-bit output. Thus, several clock cycles of input result in only one clock cycle of output.
  • the system preprocessor is designed to transform the original input data x into samples ⁇ such that small values of ⁇ ', occur much more frequently than large values.
  • the FIFO 32 is adapted to provide rate buffering when a large value is processed by temporarily storing the intermediate format data while the zero-word expander generates its all-zero outputs, allowing the output to catch up with the input.
  • the output rate can therefore keep up indefinitely with the input rate without requiring a faster internal processing clock.
  • the novel encoding method of the present invention splits the processing of incoming data streams into a pre-rate buffer process and a post-rate buffer process.
  • the separation of the processing job into these particular functions enables the use of internal rate buffering to provide full-speed input and output.
  • the pre-rate buffer process includes the conversion of the incoming data stream into the intermediate format by the FSIF packer 30. This process keeps up with the input rate.
  • the post- rate buffer process includes the conversion of the intermediate format data into the fundamental sequence coded output by the zero-word expander 34. This process may fall behind when a large value is input.
  • the FIFO 32 absorbs any short-term delays in the zero-word expander 34.
  • the encoder 14 also includes an output assembly manager 26 that assembles the code ID and reference fields from the code ED/reference FIFO 20, the FS coded data stream from the FS generator 24, and the k-split data stream from the k-split generator 22 into a final coded data set.
  • the output assembly manager 26 includes a multi-bit shifter 70, an output accumulator 72, and a timing and control unit 74.
  • the timing and control unit 74 is adapted to receive control signals from the code ID/reference FIFO 20, zero-word expander 34, and k-split FIFO 52 and in accordance therewith, generate control signals for the shifter 70 and accumulator 72.
  • the shifter 70 is a 32-bit (2L bits) register adapted to pull words from the code ID/reference FIFO 20, zero-word expander 34, or k-split FIFO 52 in accordance with the control signal from the timing and control unit 74.
  • the shifter 70 first loads the code ID and any reference data from the code ID/reference FIFO 20.
  • the first 16-bit word output from the zero-word expander 34 is loaded immediately following the code ID and reference data. Thus, if the code ID and reference data take up a total of 5 bits, they are stored in bit positions 0 through 4, and the first FS word is loaded into bit positions 5 through 20.
  • the first 16 bits are pulled by the accumulator 72 and output from the encoder 14. The remaining bits are shifted over (to start at bit position 0) and the next FS word is loaded. The first 16 bits are pulled by the accumulator 72 and the remaining bits are shifted over. This process continues until all of the FS words generated by one data block are loaded into the shifter 70. The first word from the k-split FIFO 52 is then loaded behind the last FS word. Again, the first 16 bits are pulled by the accumulator 72 and the remaining bits are shifted over. The next k-split word is then loaded into the shifter 70. This process continues until all of the k-split words generated by one data block are loaded into the shifter 70 and output by the accumulator 72. The data stream of words output by the accumulator 72 form the final coded data set y.
  • the encoder architecture of the present invention is easily implementable in multiple digital integrated circuit technologies (ASIC, FPGA, etc.). With a hardware implementation, the encoder can process data at high rates while consuming a minimum amount of power, weight, and circuit board area (low gate count).
  • the novel split-stream technique (disassembling the input stream into a fundamental sequence stream and a k-split stream, and processing the two in parallel) and the novel split-processing technique (dividing the fundamental sequence encoding into a pre- rate buffer process and a post-rate buffer process by introducing a fundamental sequence intermediate format) allow for full speed compression where the output rate indefinitely keeps up with the input rate, i.e., encoding can run as fast as the fastest internal processing clock.

Landscapes

  • Engineering & Computer Science (AREA)
  • Theoretical Computer Science (AREA)
  • Compression, Expansion, Code Conversion, And Decoders (AREA)
  • Compression Or Coding Systems Of Tv Signals (AREA)

Abstract

A data encoder (14). The novel encoder (14) includes a first circuit (24) for generating a fundamental sequence coded data stream from an incoming input data stream, a second circuit (22) for generating a k-split data stream from the incoming data stream, and a third circuit (26) for combining the fundamental sequence coded data stream and k-split data stream to form a final encoded output. The first circuit (24) includes a circuit (30) for converting the incoming input data stream into a novel intermediate format comprising a set bit word and a zero word count, and a zero-word expander (34) for converting the intermediate format to the fundamental sequence coded data stream. The first circuit (24) may also include a register (32) adapted to store the intermediate format to provide rate buffering.

Description

DATA ENCODER
BACKGROUND OF THE INVENTION
Field of the Invention:
The present invention relates to signal processing systems. More specifically, the present invention relates to data compression encoders.
Description of the Related Art:
Data compression is used in a variety of applications to encode data using fewer bits than the original representation in order to reduce the consumption of resources such as storage space or transmission bandwidth. Lossless data compression accomplishes this without any loss of information; that is, the original data can be reconstructed exactly from the encoded data.
Lossless data compression algorithms typically include two sections: a preprocessor, which transforms the input data using a statistical model into samples that can be more efficiently compressed (so that certain symbols occur more frequently than others), and an encoder, which encodes the transformed data using a scheme such that more probable symbols produce shorter output than improbable symbols. Entropy encoders encode symbols such that the code length is proportional to the probability of the symbol. More common symbols therefore use the shortest codes.
Several entropy encoding algorithms are known in the art. The Consultive Committee for Space Data Systems (CCSDS) has recommended the Rice algorithm. A Rice encoder divides an input word into two variable-length sections. One section is encoded using fundamental sequence encoding, and the other section is left alone (remains binary).
A Rice encoder is usually implemented in software running on a computer system. This implies high power consumption, weight, size, and recurring cost. Certain applications, particularly space or airborne applications, have size, weight, and power constraints that would prohibit the use of a large computer system. In addition, some applications require that the encoded data is output at a rate matching that of the incoming data. These applications require an encoder that can operate at faster processing speeds than can be achieved with conventional software implementations. A hardware approach could potentially offer faster processing speeds as well as smaller size, weight, and power consumption; however, conventional encoder architectures are either too large to realize in current digital technologies, or are too slow (i.e., output rates are slower than input rates). Hence, a need exists in the art for an improved encoder offering smaller size, weight, and power consumption, as well as faster processing speeds than conventional implementations.
SUMMARY OF THE INVENTION
The need in the art is addressed by the data encoder of the present invention. The novel encoder includes a first circuit for generating a fundamental sequence coded data stream from an incoming input data stream, a second circuit for generating a k-split data stream from the incoming data stream, and a third circuit for combining the fundamental sequence coded data stream and k-split data stream to form a final encoded output. The first circuit includes a circuit for converting the incoming input data stream into a novel intermediate format comprising a set bit word and a zero word count, and a zero-word expander for converting the intermediate format to the fundamental sequence coded data stream. The encoder may also include a register adapted to store the intermediate format to provide rate buffering to allow for full speed input and output rates.
BRIEF DESCRIPTION OF THE DRAWINGS
Fig. 1 is a simplified block diagram of a lossless data compression system designed in accordance with an illustrative embodiment of the present invention. Fig. 2 is a diagram showing an illustrative coded data set format. Fig. 3 is a simplified block diagram of an illustrative embodiment of an entropy encoder designed in accordance with the present teachings.
DESCRIPTION OF THE INVENTION
Illustrative embodiments and exemplary applications will now be described with reference to the accompanying drawings to disclose the advantageous teachings of the present invention.
While the present invention is described herein with reference to illustrative embodiments for particular applications, it should be understood that the invention is not limited thereto. Those having ordinary skill in the art and access to the teachings provided herein will recognize additional modifications, applications, and embodiments within the scope thereof and additional fields in which the present invention would be of significant utility.
Fig. 1 is a simplified block diagram of a lossless data compression system 10 designed in accordance with an illustrative embodiment of the present invention. The system 10 includes a preprocessor 12 and a novel entropy encoder 14. The preprocessor 12 receives the input data x and transforms it into preprocessed data δ suitable for the entropy encoding (typically reducing the entropy of the data stream). The entropy encoder 14 then converts the preprocessed data δ into an encoded bit sequence y.
The input data x is processed in blocks of J /i-bit words:
X={X|, X2,...Xj} [1]
where x,- is an π-bit word and n is a constant.
The preprocessor transforms the input data x into blocks of preprocessed samples δ:
e={5 ,, 52, ...A ,... 6j} [2]
where δ/ is an π-bit integer between 0 and 2"-l. Typically, the preprocessor function transforms the input data x such that the preprocessed samples δ are statistically independent and identically distributed, and . the probability that any sample δ, is a value m is a nonincreasing function of m for m = 0, 1, ... 2"— 1. Ideally, the preprocessed data stream is weighted heavily towards zero. One simple method for achieving this weighting is to make a prediction of the next value in the data stream based on the current value, and then transmit the difference between predicted and actual. For example, if the original data stream x is {100, 99, 98, 99}, the preprocessed data stream δ would be {100, -1, -1, +1}. Note that while the first value of this stream is a large number, the subsequent values are much smaller.
The preprocessor may also output a code ID and/or reference data that can be used in a decoding stage to reverse the preprocessor transform function and recover the original input data.
In an illustrative embodiment, the system 10 uses a Rice encoding algorithm. The Rice algorithm splits each preprocessed sample δ,- into two sections. The k least significant bits are split off from each sample and the remaining bits (the n-k most significant bits) are encoded using a fundamental sequence (FS) codeword. The variable k can vary for different blocks of J words (all samples within a block are encoded using the same k split). The preprocessor selects a value for k that optimizes compression for that block, and outputs a code ID field that indicates the value of A:.
Fundamental sequence encoding uses variable-length codewords, where an integer value m is represented by m zeros followed by a one. The following table shows the FS codeword for various values of the preprocessed sample δ,-:
δ/ FS Codeword
0 1 1 01
2 001
2"-l 0000...00001
(2"-l zeros)
Compression is therefore achieved when smaller values of δ, occur more frequently than larger values. The FS codewords from each of the J preprocessed words are concatenated to form a single sequence, which is output along with the removed k-split bits. The final encoded output y includes a coded data set (CDS) for each block of J input words. Fig. 2 is a diagram showing an illustrative coded data set format. Each CDS includes a code ID field that indicates the value of k, an optional reference field that includes information from the preprocessor used to reverse the preprocessor transformation, an FS field that includes the concatenated FS codewords from all J words, and a k-split field including all of the k-split least significant bits removed from the J words.
As discussed above, the entropy encoder is usually implemented in software. Certain applications would benefit from a hardware solution that can offer smaller size, weight and power consumption, as well as faster processing speeds. The present invention provides a novel entropy encoder architecture suitable for hardware implementations. The novel algorithm disassembles the uncoded input data stream into two separate streams (a fundamental sequence stream and a k-split stream) that are each partially pre-coded and partially assembled, then later processed and recombined into the final output.
Fig. 3 is a simplified block diagram of an illustrative embodiment of an entropy encoder 14 designed in accordance with the present teachings. The novel encoder 14 includes a code ID/reference FIFO (first in, first out register) 20 for storing the code ID and optional reference fields, a k-split generator 22 for computing, buffering and assembling the k-split output stream, a fundamental sequence generator 24 for computing and assembling the FS output stream, and an output assembly manager 26 for assembling the code ID and reference fields, the FS output stream, and the k-split output stream into a final coded data set. The FS generator 24 and k-split generator 22 operate in parallel to maintain full-speed input and output. The k-split generator 22 includes a k-split packer 50 and a k-split FIFO 52.
The k-split packer 50 accepts the incoming stream of uncoded data δ and packs the k split bits of each data word δ, together into a sequence of M-bit words. The k-split packer 22 includes a multi-bit shifter/masker 42 and a k-split accumulator 44. The multi-bit shifter/masker 42 is a single cycle or pipelinable shifter capable of different shift values every clock cycle. The shifter/masker 42 receives each it-bit uncoded input data word δ, and the value of k, and masks out the n—k most significant bits, keeping only the k least significant bits. The Ar-split bits are shifted and accumulated in the k-split accumulator 44 to form an M-bit word comprising the k-split bits from one or more data words packed together end-to-end. When the accumulator 44 is full, the M-bit word is sent to the k-split FIFO 52, and the next k-split bits are accumulated in the next M-bit word. Thus, when the M-bit words are strung together end-to-end, they form the final k-split data field. The FIFO 52 is large enough to store the prepacked k-split data for at least one full CDS, thus allowing the final packing and output stage to stream data out at full speed. In an illustrative embodiment, M is 16 bits; however, other word lengths can be used without departing from the scope of the present teachings.
The FS generator 24 uses a novel FS encoding technique that includes transforming the incoming stream of uncoded data into an intermediate format. The novel fundamental sequence intermediate format (FSIF) includes two fields: a pre- packed set bit word and a zero word count. The set bit word is an L-bit word with one or more bits set to '1', each set bit representing a full or partial uncoded data word. The zero word count represents the number of all-zero words to be inserted between the pre-packed set bit words. This transformation prepares the incoming data for easy conversion to full fundamental sequence coding while allowing the input stream to be accepted at full speed. The intermediate format data can be stored in a shallow FIFO to absorb any short term bursts of inefficient compression. The intermediate format data is then converted to full FS code and output at full speed.
In the illustrative embodiment, the FS generator 24 includes a FSIF packer 30, a fundamental sequence FIFO 32, and a zero-word expander 34. The FSIF packer 30 accepts the incoming stream of uncoded data δ and transforms it into the intermediate format comprised of a pre-packed set bit word and a zero word count. This intermediate format is stored in the FIFO 32. The zero-word expander 34 pulls the intermediate format data from the FIFO 32 and converts it to a fundamental sequence coded output by expanding the zero word count into L-bit all-zero words and merging them with the pre-packed set bit words.
In the illustrative embodiment, the FSEF packer 30 includes a multi-bit shifter/masker 40, a set bit calculator 42, a set bit accumulator 44, and an FSIF distributor 46. The multi-bit shifter/masker 40 receives each uncoded input data word δ,- and the value of k, and masks out the k least significant bits. The remaining n—k bit word δ't is sent to the set bit calculator 42.
The set bit calculator 42 includes an adder adapted to add the truncated data words δ'j together in a particular way and output the resulting sums using a special "one- hot" format, and logic for generating a zero word count. A one-hot format word has one and only one bit set to '1'; all other bits are set to zero. The position of the '1' defines the value of the word. For example, consider an L-bit word b={bL-i, ..., by,..., t>2, b|, bo}. A value of/ encoded in one-hot would be all zeros except for a '1 ' in the b/ position. Thus. a value of '2' would be encoded as {000...000100}.
The first sum Si output by the set bit calculator 42 is formed by adding the first truncated word δ'i to zero. The result is output by the set bit calculator 42 in one-hot format. All subsequent sums s,- are formed by adding the incoming word δ',- to the previous sum s,--y plus one. Thus:
si=δ'i
Figure imgf000009_0001
Figure imgf000009_0002
Each sum s,- is output in one-hot format to the set bit accumulator 44. The set bit calculator 42 is also adapted to output a zero word count in addition to an L-bit one- hot encoded word, hi an illustrative embodiment, L is set to 16 bits (other word lengths can be used without departing from the scope of the present teachings). When the calculator 42 tries to encode a number greater than L— 1 in one-hot format, the adder overflows. The calculator therefore includes logic for determining when the adder overflows and by how much, i.e., by how many sets of L zeros. This zero word count is output to the FSIF distributor 46.
The set bit accumulator 44 accumulates the sums output from the set bit calculator 42 to form a pre-packed set bit word with one or more bits set to T, each set bit representing a full or partial uncoded data word. The accumulator 44 starts with an L-bit word of all zeros. Each word output by the set bit calculator 42 includes one and only one bit set to ' 1 '. Each time the accumulator 44 receives a word from the calculator 42, the accumulator 44 sets one of its bits (in the same position as the T bit in the word output from the calculator 42) to 1V. This process is repeated for each word the accumulator 44 receives, until the calculator 42 overflows. The accumulator 44 thus keeps track of all the ' 1 ' bits from one or more words. When the zero word count indicates that the set bit calculator 42 has overflowed, the L-bit word in the accumulator 44 is output to the FSIF distributor 46. The accumulator 44 then resets to zero and begins to accumulate the next L-bit word. The FSIF distributor 46 is a timing and control function used to parse the set bit accumulator and set bit calculator outputs into the intermediate format. The distributor 46 is adapted to receive the zero word count from the set bit calculator 42. When the zero word count indicates that the set bit calculator 42 has overflowed, the distributor 46 pulls the L-bit word from the accumulator 44 and outputs the L-bit word and the zero word count to the FIFO 32.
The following is a short numerical example illustrating the operation of the FSIF packer 30. Consider an input sequence δ' of {2, 1, 3, 32, 2} and L=I 6 bits. On the first clock cycle, the set bit calculator 42 receives a '2' and therefore outputs a 2 encoded in one-hot format to the accumulator 44:
s, = δ', = 2 = {0000000000000100} [4]
The accumulator 44 receives the sum si and accumulates it with its previous state (all zeros), resulting in {0000000000000100}. On the second cycle, the set bit calculator 42 receives a ' 1 ' and outputs the following sum encoded in one-hot format to the accumulator 44:
S2= si+δ'2+l = 2+1+1 = 4 = {0000000000010000} [5]
The accumulator 44 accumulates the new sum s2 with its previous state, resulting in {0000000000010100}.
On the third cycle, the set bit calculator 42 receives a '3' and outputs the following sum to the accumulator 44:
S3 = s2+δ'3+l = 4+3+1 = 8 = {0000000100000000} [6] The accumulator 44 accumulates the new sum S3 with its previous state, resulting in {0000000100010100}. This accumulator output is the FS code for (reading from right to left) the sequence {2, 1, 3}. The accumulator 44 is thus forming portions of the concatenated fundamental sequence coded output. On the fourth cycle, the set bit calculator 42 receives a '32'. When it tries to add it to its previous sum, the adder overflows:
S4= s3+δ'4+l = 8+32+1 = 41 = {1 followed by 41 zeros} [7]
The set bit calculator logic therefore outputs a zero word count of '2'
(representing two sets of 16 zeros, or 32 zeros) to the distributor 46 and the 16-bit remainder {0000000100000000} to the accumulator 44. When the distributor 46 receives the zero word count indicating that the set bit calculator 42 has overflowed, it pulls the 16-bit word that was in the accumulator 44 (before the overflow). In this example, it pulls the word {0000000100010100} from the accumulator 44 and outputs the word and the zero word count to the FIFO 32. The zero word count of '2' indicates that the zero-word expander 34 should output the 16-bit word from the accumulator, followed by one 16-bit word of all zeros. (A zero word count of '3' would direct it to output the 16-bit word form the accumulator, followed by two 16-bit words of all zeros.) The accumulator 44 is then reset to zero and accumulates the new 16-bit remainder {0000000100000000} from the set bit calculator 42.
On the fifth cycle, the set bit calculator 42 receives a '2' and outputs the following sum to the accumulator 44:
S5 = s4+δ's+l = 8+2+1 = 11 - {0000100000000000} [8]
where the previous sum S4 is set to the value of the overflow remainder (s4 =8).
The accumulator 44 accumulates the new sum S5 with its previous state, resulting in {0000100100000000}. This process continues until all J words in the incoming data block δ are processed. Returning to the FS generator 30, the intermediate format data output from the distributor 46 is stored in the FIFO 32. The zero-word expander 34 pulls the intermediate format data (the pre-packed set bit word and the zero word count) from the FIFO 32 and converts it to a full, fundamental sequence coded output. The zero- word expander 34 includes a first circuit adapted to convert the zero word count to a number of L-bit words of all zeros and a second circuit adapted to output the set bit word followed by the all-zero words (if any). A zero word count of ' 1 ' corresponds with no all-zero words, so the zero-word expander 34 just outputs the pre-packed set bit word and waits to receive the next set bit word and zero word count. A zero word count of '2' corresponds with one all-zero word, so the zero-word expander 34 outputs the pre-packed set bit word followed by one L-bit word of all zeros. The sequence of words output from the zero-word expander 34 forms the concatenated FS coded output.
The FIFO 32 provides internal rate buffering to allow full-speed input and output rates. When the FSIF packer 30 processes an input word δ',- having a large value (i.e., resulting in a zero word count greater than 1), the output from the zero- word expander 34 will take more than one clock cycle: one cycle for outputting the pre-packed set bit word and an additional clock cycle for each all-zero word. On the other hand, when the FSIF packer 30 processes input words δ', having small values (i.e., much smaller than L-I), several input words are accumulated into a single L-bit output. Thus, several clock cycles of input result in only one clock cycle of output.
The system preprocessor is designed to transform the original input data x into samples δ such that small values of δ', occur much more frequently than large values.
The FIFO 32 is adapted to provide rate buffering when a large value is processed by temporarily storing the intermediate format data while the zero-word expander generates its all-zero outputs, allowing the output to catch up with the input. The output rate can therefore keep up indefinitely with the input rate without requiring a faster internal processing clock.
Thus, the novel encoding method of the present invention splits the processing of incoming data streams into a pre-rate buffer process and a post-rate buffer process. The separation of the processing job into these particular functions enables the use of internal rate buffering to provide full-speed input and output. The pre-rate buffer process includes the conversion of the incoming data stream into the intermediate format by the FSIF packer 30. This process keeps up with the input rate. The post- rate buffer process includes the conversion of the intermediate format data into the fundamental sequence coded output by the zero-word expander 34. This process may fall behind when a large value is input. The FIFO 32 absorbs any short-term delays in the zero-word expander 34.
The encoder 14 also includes an output assembly manager 26 that assembles the code ID and reference fields from the code ED/reference FIFO 20, the FS coded data stream from the FS generator 24, and the k-split data stream from the k-split generator 22 into a final coded data set. In the illustrative embodiment, the output assembly manager 26 includes a multi-bit shifter 70, an output accumulator 72, and a timing and control unit 74. The timing and control unit 74 is adapted to receive control signals from the code ID/reference FIFO 20, zero-word expander 34, and k-split FIFO 52 and in accordance therewith, generate control signals for the shifter 70 and accumulator 72.
In the illustrative embodiment, the encoder output y is a stream of L-bit words, where L= 16 bits. The shifter 70 is a 32-bit (2L bits) register adapted to pull words from the code ID/reference FIFO 20, zero-word expander 34, or k-split FIFO 52 in accordance with the control signal from the timing and control unit 74. The shifter 70 first loads the code ID and any reference data from the code ID/reference FIFO 20. The first 16-bit word output from the zero-word expander 34 is loaded immediately following the code ID and reference data. Thus, if the code ID and reference data take up a total of 5 bits, they are stored in bit positions 0 through 4, and the first FS word is loaded into bit positions 5 through 20. The first 16 bits are pulled by the accumulator 72 and output from the encoder 14. The remaining bits are shifted over (to start at bit position 0) and the next FS word is loaded. The first 16 bits are pulled by the accumulator 72 and the remaining bits are shifted over. This process continues until all of the FS words generated by one data block are loaded into the shifter 70. The first word from the k-split FIFO 52 is then loaded behind the last FS word. Again, the first 16 bits are pulled by the accumulator 72 and the remaining bits are shifted over. The next k-split word is then loaded into the shifter 70. This process continues until all of the k-split words generated by one data block are loaded into the shifter 70 and output by the accumulator 72. The data stream of words output by the accumulator 72 form the final coded data set y.
The encoder architecture of the present invention is easily implementable in multiple digital integrated circuit technologies (ASIC, FPGA, etc.). With a hardware implementation, the encoder can process data at high rates while consuming a minimum amount of power, weight, and circuit board area (low gate count). The novel split-stream technique (disassembling the input stream into a fundamental sequence stream and a k-split stream, and processing the two in parallel) and the novel split-processing technique (dividing the fundamental sequence encoding into a pre- rate buffer process and a post-rate buffer process by introducing a fundamental sequence intermediate format) allow for full speed compression where the output rate indefinitely keeps up with the input rate, i.e., encoding can run as fast as the fastest internal processing clock.
Thus, the present invention has been described herein with reference to a particular embodiment for a particular application. Those having ordinary skill in the art and access to the present teachings will recognize additional modifications, applications and embodiments within the scope thereof. For example, while the invention has been described with reference to a Rice encoder, the novel techniques described can be used with other encoding algorithms without departing from the scope of the present teachings. The encoder can also be configurable for different data widths, coded data set sizes, code selection schemes, etc.
It is therefore intended by the appended claims to cover any and all such applications, modifications and embodiments within the scope of the present invention. Accordingly,
WHAT IS CLAIMED IS:

Claims

[EUROSTYLE] CLAIMS
1. An encoder (14) characterized by: a first circuit (24) for generating a fundamental sequence coded data stream from an incoming input data stream; a second circuit (22) for generating a k-split data stream from said incoming data stream; and a third circuit (26) for combining said fundamental sequence coded data stream and k-split data stream to form a final encoded output.
2. The invention of Claim 1 wherein said second circuit (22) operates in parallel with said first circuit (24).
3. The invention of Claim 1 wherein said first circuit (24) includes a fourth circuit (30) for converting said incoming input data stream into an intermediate format comprising a set bit word and a zero word count.
4. The invention of Claim 3 wherein said first circuit (24) further includes a zero-word expander (34) for converting said intermediate format to a fundamental sequence coded data stream.
5. The invention of Claim 4 wherein said first circuit (24) further includes a register (32) for storing said intermediate format to provide rate buffering.
6. The invention of Claim 5 wherein said fourth circuit (30) includes a multi-bit shifter/masker (40) for receiving said incoming input data stream and an integer value k, and masking out k bits from each word of said input stream to output a stream of truncated input data words.
7. The invention of Claim 6 wherein said fourth circuit (30) further includes a set bit calculator (42) for adding said truncated input data words and outputting a resulting sum.
8. The invention of Claim 7 wherein said sum is output as an L-bit word encoded in one-hot format.
9. The invention of Claim 8 wherein said set bit calculator (42) also outputs a zero word count indicating when a sum is larger than can be encoded as an L-bit word and by how many sets of L-bit all-zero words.
10. The invention of Claim 9 wherein a first sum is equal to a first truncated input word and subsequent sums are equal to a previous sum plus an incoming truncated input word plus one.
11. The invention of Claim 10 wherein said fourth circuit (30) further includes an accumulator (44) for accumulating said sums to form a set bit word.
PCT/US2007/017893 2006-08-17 2007-08-14 Data encoder Ceased WO2008021304A2 (en)

Priority Applications (2)

Application Number Priority Date Filing Date Title
EP07811287.7A EP2055007B1 (en) 2006-08-17 2007-08-14 Data encoder
IL197045A IL197045A (en) 2006-08-17 2009-02-15 Data encoder

Applications Claiming Priority (2)

Application Number Priority Date Filing Date Title
US11/505,749 2006-08-17
US11/505,749 US7504970B2 (en) 2006-08-17 2006-08-17 Data encoder

Publications (2)

Publication Number Publication Date
WO2008021304A2 true WO2008021304A2 (en) 2008-02-21
WO2008021304A3 WO2008021304A3 (en) 2008-10-30

Family

ID=39082668

Family Applications (1)

Application Number Title Priority Date Filing Date
PCT/US2007/017893 Ceased WO2008021304A2 (en) 2006-08-17 2007-08-14 Data encoder

Country Status (5)

Country Link
US (1) US7504970B2 (en)
EP (2) EP2055007B1 (en)
IL (1) IL197045A (en)
TW (1) TWI378652B (en)
WO (1) WO2008021304A2 (en)

Families Citing this family (7)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
KR101487190B1 (en) * 2008-09-11 2015-01-28 삼성전자주식회사 Flash memory integrated circuit having a compression codec
US8228911B2 (en) * 2008-09-19 2012-07-24 Honeywell International Inc. Enhanced data link communication over iridium
US7719441B1 (en) * 2009-01-05 2010-05-18 Honeywell International Inc. System and method for transferring bit-oriented data over an ACARS character-oriented data link
US8401600B1 (en) 2010-08-02 2013-03-19 Hypres, Inc. Superconducting multi-bit digital mixer
KR102100408B1 (en) * 2014-03-04 2020-04-13 삼성전자주식회사 Encoder resistant to power analysis attack and encoding method thereof
EP3935581A4 (en) 2019-03-04 2022-11-30 Iocurrents, Inc. Data compression and communication using machine learning
US11601136B2 (en) * 2021-06-30 2023-03-07 Bank Of America Corporation System for electronic data compression by automated time-dependent compression algorithm

Family Cites Families (14)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US4506252A (en) * 1981-05-05 1985-03-19 Sperry Corporation Ternary data encoding system
US5550540A (en) * 1992-11-12 1996-08-27 Internatioal Business Machines Corporation Distributed coding and prediction by use of contexts
US5663725A (en) * 1995-11-08 1997-09-02 Industrial Technology Research Institute VLC decoder with sign bit masking
ATE189353T1 (en) * 1996-04-18 2000-02-15 Nokia Mobile Phones Ltd VIDEO DATA ENCODER AND DECODER
CA2296060C (en) * 1997-07-09 2004-06-29 Quvis, Inc. Apparatus and method for entropy coding
US6757334B1 (en) * 1998-08-10 2004-06-29 Kamilo Feher Bit rate agile third-generation wireless CDMA, GSM, TDMA and OFDM system
US6885319B2 (en) * 1999-01-29 2005-04-26 Quickshift, Inc. System and method for generating optimally compressed data from a plurality of data compression/decompression engines implementing different data compression algorithms
US6567127B1 (en) * 1999-10-08 2003-05-20 Ati International Srl Method and apparatus for enhanced video encoding
US6480125B2 (en) * 2000-06-09 2002-11-12 Seagate Technology Llc Method and apparatus for efficient encoding of large data words at high code rates
US6614369B1 (en) * 2002-03-05 2003-09-02 International Business Machines Corporation DC balanced 7B/8B, 9B/10B, and partitioned DC balanced 12B/14B, 17B/20B, and 16B/18B transmission codes
US7082168B2 (en) * 2002-05-21 2006-07-25 Coffey John T Methods and apparatus for self-inverting turbo code interleaving with high separation and dispersion
US6781435B1 (en) * 2003-02-03 2004-08-24 Hypres, Inc. Apparatus and method for converting a multi-bit signal to a serial pulse stream
US20060197689A1 (en) * 2005-03-02 2006-09-07 Regents Of The University Of Minnesota Parallelized binary arithmetic coding
US7262722B1 (en) * 2006-06-26 2007-08-28 Intel Corporation Hardware-based CABAC decoder with parallel binary arithmetic decoding

Non-Patent Citations (1)

* Cited by examiner, † Cited by third party
Title
None

Also Published As

Publication number Publication date
TWI378652B (en) 2012-12-01
EP2055007B1 (en) 2017-11-01
IL197045A (en) 2013-06-27
EP2367290A1 (en) 2011-09-21
EP2055007A2 (en) 2009-05-06
IL197045A0 (en) 2009-11-18
TW200822578A (en) 2008-05-16
WO2008021304A3 (en) 2008-10-30
US20080055121A1 (en) 2008-03-06
US7504970B2 (en) 2009-03-17

Similar Documents

Publication Publication Date Title
JP3136796B2 (en) Variable length code decoder
US5912636A (en) Apparatus and method for performing m-ary finite state machine entropy coding
CN113810057B (en) Method, apparatus and system for semantic value data compression and decompression
JP4422833B2 (en) Decoder and method for decoding variable length codeword
EP2055007B1 (en) Data encoder
US20030085822A1 (en) High performance memory efficient variable-length coding decoder
JP3294026B2 (en) High-speed variable-length decoding device
US6339386B1 (en) Variable length coder of a video coder
EP1958450A2 (en) Decoding data
US5663726A (en) High speed variable-length decoder arrangement with reduced memory requirements for tag stream buffering
US5666116A (en) High speed variable-length decoder arrangement
KR0152032B1 (en) Variable long decoder for video signal
KR19980702418A (en) Variable length decoder
KR100499966B1 (en) Signal processor
Mehboob et al. High speed lossless data compression architecture
JP3863652B2 (en) Variable length code alignment device
Pasumarthi et al. Ehtc: An enhanced huffman tree coding algorithm and its fpga implementation
Biasizzo et al. A multi–alphabet arithmetic coding hardware implementation for small fpga devices
KR100275267B1 (en) High speed variable length code decoding device
Dey Iterative Data Compression with Calculated Codes-IEEE_1col
KR0125125B1 (en) High speed variable length code decoding device
KR960011111B1 (en) Variable length decoder of digital image signal
Doshi et al. Improved performance of arithmetic coding by extracting multiple bits at a time
KR0125126B1 (en) High speed variable length code decoding device
Xue et al. Efficient VLSI Implementation of a VLC Decoder for Golomb-Rice Code using Alternating Coding

Legal Events

Date Code Title Description
WWE Wipo information: entry into national phase

Ref document number: 197045

Country of ref document: IL

NENP Non-entry into the national phase

Ref country code: DE

REEP Request for entry into the european phase

Ref document number: 2007811287

Country of ref document: EP

WWE Wipo information: entry into national phase

Ref document number: 2007811287

Country of ref document: EP

NENP Non-entry into the national phase

Ref country code: RU