WO2003083859A2 - Watermark time scale searching - Google Patents

Watermark time scale searching Download PDF

Info

Publication number
WO2003083859A2
WO2003083859A2 PCT/IB2003/000794 IB0300794W WO03083859A2 WO 2003083859 A2 WO2003083859 A2 WO 2003083859A2 IB 0300794 W IB0300794 W IB 0300794W WO 03083859 A2 WO03083859 A2 WO 03083859A2
Authority
WO
WIPO (PCT)
Prior art keywords
ofthe
sequence
watermark
estimate
signal
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Ceased
Application number
PCT/IB2003/000794
Other languages
French (fr)
Other versions
WO2003083859A3 (en
Inventor
Aweke N. Lemma
Leon M. Van De Kerkhof
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Koninklijke Philips NV
Original Assignee
Koninklijke Philips Electronics NV
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Koninklijke Philips Electronics NV filed Critical Koninklijke Philips Electronics NV
Priority to KR10-2004-7015275A priority Critical patent/KR20040097227A/en
Priority to DE60308667T priority patent/DE60308667T2/en
Priority to JP2003581193A priority patent/JP4302533B2/en
Priority to EP03710065A priority patent/EP1493145B1/en
Priority to US10/509,411 priority patent/US7266466B2/en
Priority to AU2003214489A priority patent/AU2003214489A1/en
Publication of WO2003083859A2 publication Critical patent/WO2003083859A2/en
Publication of WO2003083859A3 publication Critical patent/WO2003083859A3/en
Anticipated expiration legal-status Critical
Ceased legal-status Critical Current

Links

Classifications

    • GPHYSICS
    • G11INFORMATION STORAGE
    • G11BINFORMATION STORAGE BASED ON RELATIVE MOVEMENT BETWEEN RECORD CARRIER AND TRANSDUCER
    • G11B20/00Signal processing not specific to the method of recording or reproducing; Circuits therefor
    • G11B20/10Digital recording or reproducing
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L19/018Audio watermarking, i.e. embedding inaudible data in the audio signal
    • GPHYSICS
    • G01MEASURING; TESTING
    • G01LMEASURING FORCE, STRESS, TORQUE, WORK, MECHANICAL POWER, MECHANICAL EFFICIENCY, OR FLUID PRESSURE
    • G01L19/00Details of, or accessories for, apparatus for measuring steady or quasi-steady pressure of a fluent medium insofar as such details or accessories are not special to particular types of pressure gauges
    • GPHYSICS
    • G11INFORMATION STORAGE
    • G11BINFORMATION STORAGE BASED ON RELATIVE MOVEMENT BETWEEN RECORD CARRIER AND TRANSDUCER
    • G11B20/00Signal processing not specific to the method of recording or reproducing; Circuits therefor
    • GPHYSICS
    • G11INFORMATION STORAGE
    • G11BINFORMATION STORAGE BASED ON RELATIVE MOVEMENT BETWEEN RECORD CARRIER AND TRANSDUCER
    • G11B20/00Signal processing not specific to the method of recording or reproducing; Circuits therefor
    • G11B20/00086Circuits for prevention of unauthorised reproduction or copying, e.g. piracy
    • G11B20/00884Circuits for prevention of unauthorised reproduction or copying, e.g. piracy involving a watermark, i.e. a barely perceptible transformation of the original data which can nevertheless be recognised by an algorithm
    • G11B20/00891Circuits for prevention of unauthorised reproduction or copying, e.g. piracy involving a watermark, i.e. a barely perceptible transformation of the original data which can nevertheless be recognised by an algorithm embedded in audio data

Definitions

  • the present invention relates to apparatus and methods for decoding information that has been embedded in information signals, such as audio, video or data signals.
  • Watermarking of information signals is a technique for the transmission of additional data along with the information signal.
  • watermarking techniques can be used to embed copyright and copy control information into audio signals.
  • a watermarking scheme The main requirement of a watermarking scheme is that it is not observable (i.e. in the case of an audio signal, it is inaudible) whilst being robust to attacks to remove the watermark from the signal (e.g. removing the watermark will damage the signal). It will be appreciated that the robustness of a watermark will normally be a trade off against the quality ofthe signal in which the watermark is embedded. For instance, if a watermark is strongly embedded into an audio signal (and is thus difficult to remove) then it is likely that the quality ofthe audio signal will be reduced. In digital devices, it is typically assumed that there exists up to a 1% drift in sampling (clock) frequency.
  • this drift is normally manifested as a stretch or shrink in the time domain signal (i.e. a linear time scale change).
  • a watermark embedded in the time domain e.g. in an audio signal
  • this time stretch or shrink will be affected by this time stretch or shrink as well, which can make watermark detection very difficult or even impossible.
  • any linear time scale change within the signal is resolved by repeatedly running the watermark detection (including repeating the extraction ofthe watermark from the host signal) for the different possible time scales, until all the possible time scales are exhausted, or detection is achieved.
  • Performing such searches over the possible time scaling ranges requires a large computational overhead, and is thus costly in terms of both hardware and computational time. Consequently, real time implementation of a watermark detector utilizing such a time scale search technique is not feasible.
  • the present invention provides a method of compensating for a linear time scale change in a received signal, the signal being modified by a sequence of symbols in the time domain, the method comprising the steps of: (a) extracting an initial estimate ofthe sequence of symbols from said received signal;(b) forming an estimate of a correctly time scaled sequence ofthe symbols by interpolating the values of said initial estimate.
  • step (b) is repeated so as to provide a range of estimates corresponding to different time scalings.
  • said interpolation is at least one of zeroth order interpolation, linear interpolation, quadratic interpolation and cubic interpolation.
  • the method further comprises the step of processing each estimate as though it were the correctly time scaled sequence ofthe symbols, so as to determine which estimate is the best estimate.
  • the method further comprises the steps of correlating each of said estimates with a reference corresponding to said sequence of symbols; and taking the estimate with the maximum correlation peak as the best estimate.
  • said initial estimate ofthe sequence of symbols is stored in a buffer.
  • said buffer is of total length M, the total bf siskin ..searghjs/
  • N ⁇ — (77 max - 77 min ) where ⁇ m i n , ⁇ mzx correspond respectively to the mimmum
  • said initial estimates ofthe sequence of symbols comprises a sequence of N b estimates for each symbol, each ofthe N b estimates corresponding to a different time offset of a symbol.
  • the scale search in the next detection window is adapted based on the information acquired during the current detection window.
  • the scale space is searched using an optimal searching algorithm.
  • the searching algorithm is the grid refinement algorithm.
  • the present invention provides a computer program arranged to perform the method as described above.
  • the present invention provides a record carrier comprising the computer program, and a method of making available for downloading the computer program.
  • the present invention provides an apparatus arranged to compensate for a linear time scale change in a received signal, the signal being modified by a sequence of symbols in the time domain, the apparatus comprising: an extractor arranged to extract an initial estimate ofthe sequence of symbols from said received signal; and an interpolator arranged to form an estimate of a correctly time scaled sequence ofthe symbols by interpolating the values of said initial estimate.
  • the apparatus further comprises a buffer arranged to store one or more of said estimates.
  • the present invention provides a decoder comprising the apparatus as described above.
  • FIG. 1 is a diagram illustrating a watermark embedding apparatus
  • Figure 2 shows a signal portion extraction filter H
  • Figures 3 a and 3b show respectively the typical amplitude and phase responses as a function of frequency ofthe filter ⁇ shown in Fig. 2;
  • Figure 4 shows the payload embedding and watermark conditioning stage of the apparatus shown in Fig. 1 ;
  • Figure 5 is a diagram illustrating the details ofthe watermark conditioning apparatus H e of Fig. 4, including charts ofthe associated signals at each stage;
  • Figure 6a and 6b show two preferred alternative window shaping functions s(n) in the form of respectively a raised cosine function and a bi-phase function;
  • Figures 7a and 7b show respectively the frequency spectra for a watermark sequence conditioned with a raised cosine and a bi-phase shaping window function
  • Figure 8 is a diagram illustrating a watermark detector in accordance with an embodiment of the present invention.
  • Figure 9 diagrammatically shows the whitening filter H w of Fig. 8, for use in conjunction with a raised cosine shaping window function
  • Figure 10 diagrammatically shows the whitening filter H w of Fig. 8, for use in conjunction with a bi-phase window shaping function;
  • Figure 11 shows details ofthe watermark symbol extraction and buffering processes in accordance with an embodiment ofthe present invention;
  • Figure 12 illustrates a sequence in which estimates of watermark symbols are collected from four buffers when there is no time scale modification
  • Figures 13a and 13b illustrate the different sequences, according to an embodiment ofthe present invention, in which estimates of watermark symbols can be collected from four buffers when there is respectively a time stretch and a time shrink time scale modification;
  • Figure 14 shows an example of an efficient scale search technique based on the concept of grid refinement
  • Figure 15 shows a typical shape ofthe correlation function output from the correlator ofthe watermark detector shown in Fig. 8.
  • Fig. 1 shows a block diagram ofthe apparatus required to perform the digital signal processing for embedding a multi-bit payload watermark w into a host signal x.
  • a host signal x is provided at an input 12 ofthe apparatus.
  • the host signal x is passed in the direction of output 14 via the adder 22.
  • a replica ofthe host signal x (input 8) is split off in the direction ofthe multiplier 18, for carrying the watermark information.
  • the watermark signal w c is obtained from the payload embedder and watermark conditioning apparatus 6, and derived from a reference finite length random sequence w s input to the payload embedder and watermark conditioning apparatus.
  • the multiplier 18 is utilized to calculate the product ofthe watermark signal w c and the replica audio signal x.
  • the resulting product, W c X is then passed via a gain controller 24 to the adder 22.
  • the gain controller 24 is used to amplify or attenuate the signal by a gain factor ⁇ .
  • the gain factor ⁇ controls the trade off between the audibility and the robustness ofthe watermark. It may be a constant, or variable in at least one of time, frequency and space.
  • the apparatus in Fig. 1 shows that, when ⁇ is variable, it can be automatically adapted via a signal analyzing unit 26 based upon the properties ofthe host signal x.
  • the gain ⁇ is automatically adapted, so as to minimize the impact on the signal quality, according to a properly chosen perceptibility cost- function, such as a psycho- acoustic model ofthe human auditory system (HAS) in case of an audio signal.
  • HAS human auditory system
  • an audio watermark is utilized, by way of example only, to describe this embodiment ofthe present invention.
  • the watermark w c is chosen such that when multiplied with x, it predominantly modifies the short time envelope of x.
  • Fig. 2 shows one preferred embodiment in which the input 8 to the multiplier
  • the filter H is a linear phase band pass filter characterized by its lower cut-off frequency.//, and upper cut-off frequency fy.
  • the filter H has a linear phase response with respect to frequency/ within the pass- band (BW).
  • BW pass- band
  • x b and x b are the in-band and out-of-band components ofthe host signal respectively.
  • the signals x b and x b are in phase. This is achieved by appropriately compensating for the phase distortion produced by filter H.
  • the distortion is a simple time delay.
  • Fig. 4 the details ofthe payload embedder and watermark conditioning unit 6 is shown.
  • the initial reference random sequence w s is converted into a multi-bit watermark signal w c .
  • a finite length, preferably zero mean and uniformly distributed random sequence w s from now on also referred to as the watermark seed signal, is generated using a random number generator with an initial seed S.
  • this initial seed S is known to both the embedder and the detector, such that a copy ofthe watermark signal can be generated at the detector for comparison purposes.
  • This results in the sequence of length L w w s [k] e [-1,1], for k 0,l,2 L w -1 (4)
  • the seed can be transmitted to the detector via an alternate channel or can be derived from the received signal using some predetermined protocol.
  • the sequence w s is circularly shifted by the amounts dj and d using the circularly shifting unit 30 to obtain the random sequences W JJ and w d2 respectively.
  • these two sequences (waj and vtt ⁇ ) are effectively a first sequence and a second sequence, with the second sequence being circularly shifted with respect to the first.
  • each sequence is then converted into a periodic, slowly varying narrow-band signal w t of length L W T S by the watermark conditioning circuit 20 shown in Fig. 4.
  • the slowly varying narrow-band signals wj and w 2 are added with a relative delay T r (where T r ⁇ T s ) to give the multi-bit payload watermark signal w c . This is achieved by first delaying the signal w 2 by the amount T r using delaying unit 45 and subsequently by adding it to wj with the adding unit 50.
  • Fig. 5 shows the watermark conditioning apparatus 20 used in the payload embedder and watermark conditioning apparatus 6 in more detail.
  • the watermark seed signal w s is input to the conditioning apparatus 20.
  • Chart 181 illustrates one ofthe sequences w ⁇ as a sequence of values of random numbers between +1 and -l s with the sequence being of length L w .
  • the sample repeater repeats each value within the watermark seed signal sequence T s times, so as to generate a rectangular pulse train signal.
  • T s is referred to as the watermark symbol period and represents the span ofthe watermark symbol in the audio signal.
  • Chart 183 shows the results ofthe signal illustrated in chart 181 once it has passed through the sample repeater 180.
  • a window shaping function s[n] such as a raised cosine window, is then applied to convert the rectangular pulse functions derived from w ⁇ and vit ⁇ into slowly varying watermark sequence functions wjfnj and w 2 [n] respectively.
  • Chart 184 shows a typical raised cosine window shaping function, which is also of span T s .
  • the generated watermark sequences wj[n] and w 2 [n] are then added up with a relative delay T r (where T r ⁇ T s ) to give the multi-bit payload watermark signal w c [n] i.e.,
  • T r The value of T r is chosen such that the zero crossings of w t match the maximum amphtude points of w and vice-versa.
  • T r TJ2
  • T r T 4
  • other values of T r are possible.
  • pL ' is an estimate ofthe circular shift pL between W d i and wj 2 , which is part ofthe payload, and is defined as
  • extra information can be encoded by changing the relative signs ofthe embedded watermarks.
  • the payload is immune to relative offset between the embedder and the detector, and also to possible time scale modifications.
  • the window shaping function has been identified as one ofthe main parameters that controls the robustness and audibility behavior ofthe present watermarking scheme.
  • two examples of possible window shaping functions are herein described - a raised cosine function and a bi-phase function. It is preferable to use a bi-phase window function instead of a raised cosine window function, so as to obtain a quasi DC-free watermark signal.
  • bi-phase window offers superior audibility performance for the same robustness or, conversely, it allows a better robustness for the same audibility quality.
  • a bi-phase function could be utilized as a window shaping function for other watermarking schemes. In other words, a bi-phase function could be applied to reduce the DC component of signals (such as a watermark) that are to be incorporated into another signal.
  • Fig. 8 shows a block diagram of a watermark detector (200, 300, 400).
  • the detector consists of three major stages: (a) the watermark symbol extraction stage (200), (b) the buffering and interpolation stage (300), and (c) the correlation and decision stage (400).
  • the received watermarked signal y '[n] is processed to generate multiple (N b ) estimates ofthe watermarked sequence. These estimates ofthe watermark sequence are required to resolve time offset that may exist between the embedder and the detector, so that the watermark detector can synchronize to the watermark sequence inserted in the host signal.
  • these estimates are demultiplexed into N b separate buffers, and an interpolation is applied to each buffer to resolve time scale modifications that may have occurred, e.g. a drift in sampling (clock) frequency may have resulted in a stretch or shrink in the time domain signal (i.e. the watermark may have been stretched or shrunk).
  • a drift in sampling (clock) frequency may have resulted in a stretch or shrink in the time domain signal (i.e. the watermark may have been stretched or shrunk).
  • the content of each buffer is correlated with the reference watermark and the maximum correlation peaks are compared against a threshold to determine the likelihood of whether the watermark is indeed embedded within the received signal y'[n].
  • each watermark symbol to be detected can be constructed by taking the average of several estimates of said symbol. This averaging process is referred to as smoothing, and the number of times the averaging is done is referred to as the smoothing factor Sf.
  • T s the symbol period and L w the number of symbols within the watermark sequence.
  • the incoming watermark signal y '[n] is input to the optional signal conditioning filter H b (210).
  • This filter 210 is typically a band pass filter and has the same behavior as the corresponding filter (H, 15) shown in Fig. 2.
  • the output ofthe filter Hi, is y'bfnj and, assuming linearity within the transmission medium, it follows from equations (1) and (3):
  • H b in the detector can also be omitted, or it can still be included to improve the detection performance. If Hb is omitted, then y b in equation (10) is replaced with y. The rest ofthe processing is the same.
  • the output y ' [nJ ofthe filter H b is provided as an input to a frame divider 220, which divides the audio signal into frames of length T s i.e. into y ' b .mfnj, with the energy calculating unit 230 then being used to calculate the energy corresponding to each ofthe framed signals as per equation (12).
  • the output of this energy calculation unit 230 is then provided as an input to the whitening stage H w (240) which performs the function shown in equation 13 so as to provide an output w e [m] .
  • Alternative implementations (240A, 240B) of this whitening stage are illustrated in Figs. 9 and 10.
  • the denominator of equation 13 contains a term that requires knowledge ofthe host (original) signal x. As the signal x is not available to the detector, it means that in order to calculate w e [m] then the denominator of equation 13 must be estimated. Below is described how such an estimation can be achieved for the two described window shaping functions (the raised cosine window shaping function and the biphase window shaping function), but it will equally be appreciated that the teaching could be extended to other window shaping functions.
  • equation 13 may be approximated by:
  • such a whitening filter H w (240A) comprises an input 242A for receiving the signal EfmJ. A portion of this signal is then passed through the low pass filter 247 A to produce a low pass filtered energy signal EipfmJ, which in turn is provided as an input to the calculation stage 248 A along with the function EfmJ. The calculation stage 248 A then divides EfmJ by EipfmJ to calculate the extracted watermark symbol w e fmj.
  • each audio frame is first sub-divided into two halves.
  • the energy functions corresponding to the first and second half-frames are hence given by
  • the original audio envelope can be approximated as the mean of EjfmJ and E 2 m/. Further, the instantaneous modulation value can be taken as the difference between these two functions.
  • the watermark w e fmj can be approximated by:
  • the whitening filter H w (240B) in Fig. 8 for a bi-phase window shaping function can be realized as shown in Fig. 10.
  • Inputs 242B and 243B respectively receive the energy functions ofthe first and second half frames EifmJ and E 2 .
  • Each energy function is then split up into two, and provided to adders 245B and 246B which respectively calculate EifmJ - E 2 [m], and EifmJ + E ⁇ fmJ.
  • Both of these calculated functions are then passed to the calculating unit 248B which divides the value from adder 245B by the value from 246B so as to calculate w e [r ⁇ /, containing N b time-multiplexed estimates ofthe embedded watermark sequences, in accordance with equation 17.
  • This output w e fmj is then passed to the buffering and interpolation stage 300 (Fig. 8), where the signal is de-multiplexed by a de-multiplexer 310, buffered in buffers 320 of length L b , so as to resolve a lack of synchronism between the embedder and the detector, and interpolated within the interpolation unit 330 so as to compensate for a time scale modification between the embedder and the detector.
  • Fig. 11 illustrates the process carried out by the buffering and interpolation stage 300 to resolve the offset issue.
  • the example described illustrates the process for resolving offset when a raised cosine window shaping function has been employed in the watermark embedding process.
  • the same technique is applicable when the bi-phase window shaping function has been used.
  • the incoming audio signal stream y ' fnJ is separated into preferably overlapping frames 302 of effective length T s by the frame divider 220.
  • each frame is divided into N b sub-frames (304a, 304b,...,304x), and the above computations (equations (12) to (17)) are applied on a sub-frame basis.
  • each sub-frame overlaps with an adjacent sub-frame.
  • the main frames are preferably longer than the symbol period T s so as to allow inter-frame overlap as shown in Fig. 11.
  • the energy ofthe audio is then computed for each sub-frame by the whitening stage 240, and the resulting values are de-multiplexed into the N b buffers 320 by the demultiplexer 310.
  • Each one(i? ⁇ B2, .... Bm) ofthe buffers 320 will thus contain a sequence of values, with the first buffer Bj containing a sequence of values corresponding to the first sub- frame within each frame, the second buffer B2 containing a sequence of values corresponding to the second sub-frame within each frame etc.
  • L b is the buffer length
  • each buffer thus contains an estimate ofthe symbol sequence, the estimates corresponding to the sequences having different time offsets.
  • the sub-frame best aligned with the center ofthe frame i.e. the best estimate ofthe correctly aligned frame
  • the sequence with the maximum correlation peak value is chosen as the best estimate ofthe correctly aligned frame.
  • the corresponding confidence level is used to determine the truth-value ofthe detection.
  • the correlation process is halted once an estimated watermark sequence with a correlation peak above the defined threshold has been found.
  • the length of each buffer is between 3 to 4 times the watermark sequence length L w , and is thus typically of length between 2048 and 8192 symbols, and Na is typically within the range of 2 to 8.
  • the buffer is normally 3 to 4 times that ofthe watermark sequence so that each watermark symbol can be constructed by taking the averages of several estimates of said symbol.
  • This averaging process is referred to as smoothing, and the number of times the averaging is done is referred to as the smoothing factor Sf.
  • the smoothing factor Sf is such that:
  • the detector ref nes ⁇ the parameters used in the offset search based upon the results of a previous search step. For instance, if a first series of estimates shows that the results stored in buffer B 3 provide the best estimate ofthe information signal, then the next offset search (either on the same received signal, or on the signal received during the next detection window) is refined by shifting the position ofthe sub-frames towards the position of the best estimate sub-frame. The estimates of the sequence having zero offset can thus be iteratively improved.
  • ⁇ b estimates ofthe watermark sequence are constructed by collecting the symbols stored in the ⁇ b buffers separately.
  • Fig. 12 illustrates four buffers (Bl, B2, B3, B4), each buffer shown as a row of boxes, with each box within a row indicating a separate location within the respective buffer.
  • the sequences w ⁇ , W12, w ⁇ 3 , w 1 are respective estimates ofthe watermark sequence.
  • each estimate (w ⁇ , w ⁇ , w I3 , w ] ) represents an estimate ofthe watermark sequence with different time offset.
  • each estimate (that is passed to the correlator 410) is formed by sequentially collecting the entries from each buffer. For example, the first value in sequence ⁇ (w ⁇ [1]) is collected from the first location of Bl, the second (w ⁇ [2]) from the second location of Bl etc, with the final value (w ⁇ [Lb]) being collected from the final location ofthe buffer.
  • the arrows which connect each box in a row to the neighboring box, show the direction in which values ofthe sequence estimates are collected from the buffer locations. It will also be appreciated that, whilst only eleven buffer locations are shown for each buffer, the size ofthe buffers in practice is likely to be significantly larger than this.
  • the length of each buffer is typically between 2048 and 8192 locations, with the number of buffers typically being between 2 and 8.
  • the actual buffer lengths are set to (l+
  • such a search is performed by systematically combining the extracted watermark sequence estimates (w e [m]), preferably by systematically combining
  • Such time scale searches can be performed by utilizing any order of interpolation.
  • two orders of interpolation will be described - the first order (linear) interpolation and the zero order interpolation.
  • this technique can be extended to higher orders of interpolation e.g. quadratic and cubic interpolation.
  • estimates ofthe time scaled watermark sequence are provided by applying linear interpolation to the previously extracted estimates ofthe watermark sequence.
  • the intermediate values w e [k] generated by the symbol extraction step shown in Fig. 8 are sequentially stored in a single buffer of length M in place ofthe N b buffers.
  • the N b buffers are multiplexed into a single buffer of length where L w and sr are as defined earlier.
  • Let the so stretched sequence be represented by w /
  • w D represents discrete samples of an otherwise continuous function. During time scale modification, these discrete points are either pushed towards each other or stretched out. This in turn is translated to re-sampling of the watermark function.
  • re-sampling is realized via a linear interpolation technique. That is, given the watermark sequence ...,M, an interpolated watermark sequence wjfmj is generated as
  • wo.bfkj be the pre-interpolation sequence stored in the ⁇ -th buffer, and q pk e ⁇ ], ...sjL w ⁇ and r p t e ⁇ l, ...N b J be defined as
  • equation (22) estimates of the time scaled watermark sequence are provided by applying zero order interpolation to the previously extracted estimates ofthe watermark sequence.
  • the interpolation function can be written as
  • FIG. 13a shows how the different estimates ofthe correct watermark sequences (w ⁇ , w ⁇ , W I3 , W 1 ) are extracted from the buffers for a time stretch
  • Fig. 13b shows similar information for a time shrink.
  • each row of boxes represents a respective buffer, with each box representing a location within each buffer.
  • the arrows indicate the order in which the buffer contents are collected from the estimates ofthe watermark sequences.
  • the watermark symbol combining stage tracks the size ofthe drift.
  • N b is the number of buffers i.e. the number of consecutive symbols that represent a single watermark symbol
  • the symbol collection sequence from the buffers is adjusted to provide the next best estimate ofthe symbol from the buffers.
  • the buffer counters are incremented or decremented (depending on drift direction), and a circular rotation ofthe buffer pointer for each watermark sequence estimation (wn, w I2 , w> / j, wu) is performed.
  • k be the buffer entry counter, where £ is an integer representing each location within each buffer i.e.
  • is positive (time stretch)
  • the counter for the first buffer is incremented.
  • the ordering ofthe buffers is also circularly shifted (i.e. the watermark sequence estimate w u previously being taken from buffer one will now been taken from buffer four, the estimate from buffer two will now be taken from buffer one, the estimate from three will now be taken from buffer two, and the estimate from buffer four will now be taken from buffer three).
  • a similar circular shift is also performed on the buffer counter k. This is shown diagrammatically in Fig. 13 a. If ⁇ is negative (time stretch), the counter for the first buffer is incremented, and the ordering ofthe buffers is circularly shifted (i.e.
  • the time scaled watermark sequence has been estimated by selecting those values from the original, non time scaled watermark sequence estimates that would most closely correspond to the temporal positions ofthe time scaled watermark sequence.
  • Such a technique efficiently resolves the problems of estimating correctly time scaled watermarks, with minimal cost in terms of computational overhead.
  • Such estimates ofthe time scaled watermark sequence will then be passed to the correlator (410), so as to determine whether the predicted time shift ⁇ accurately represents the time shift ofthe received signal i.e. do the estimates provided to the correlator provide good correlation peaks.
  • the time scale search will be repeated for a different estimated value i.e. a different value of ⁇ . Due to possible time scale modification, the detection truth-value (whether or not the signal includes a watermark) is determined only after the appropriate scale search has been conducted.
  • be the scale search step size and let us assume that we want the watermark to survive all the scale modifications in the interval [/ / "min, max]- The total number of visited scales is then given by T ⁇ max ⁇ ?7min .. .. A ⁇ (24)
  • the scale search is adapted such that information acquired during detection is utilized to plan an optimum search in the subsequent detection windows. For example, the scale search in the next detection window is started around the current optimum scale.
  • FIG. 14 An alternative embodiment illustrated in Fig. 14 provides a method for efficient walk through the scale space by grid refinement.
  • the most straightforward solution is a linear search from the minimum scale towards the maximum scale by adding up an incremental step. Assuming correlation, and thus confidence level, does not change abruptly from one scale to the next, one can considerably reduce the amount of scales visited during the search by reducing the space granularity.
  • the algorithm starts at scale zero and is repeated until a minimum granularity is reached or the watermark is detected (i.e., a local maximum for the confidence level is found) and/or the confidence level exceeds a predetermined threshold.
  • a random or linear search around this scale may suffice.
  • outputs (WD I , W D2 , ... w DNb ) from the buffering stage are passed to the interpolation stage and, after interpolation, the outputs (wn, wn, ... w/w,) of this stage, which are needed to resolve a possible time scale modification in the watermarked signal, are passed to the correlation and decision stage. All ofthe estimates (wn, wn,- wmb) ofthe watermark corresponding to the different possible offset values are passed to the correlation and decision stage 400.
  • the correlator 410 calculates the correlation of each estimate ...,N b with respect to the reference watermark sequence w c fkj. Each respective correlation output corresponding to each estimate is then applied to the maximum detection unit 420 which determines which two estimates provided the maximum correlation peak values. These estimates are chosen as the ones that best fit the circularly shifted versions W d i and Wd 2 ofthe reference watermark.
  • the correlation values for these estimated sequences are passed to the threshold detector and payload extractor unit 430.
  • the reference watermark sequence w s used within the detector corresponds to
  • the detector can calculate the same random number sequence using the same random number generation algorithm and the same initial seed S so as to determine the watermark signal.
  • the watermark signal originally applied in the embedder and utilized by the detector as a reference could simply be any predetermined sequence.
  • Fig. 15 shows a typical shape of a correlation function as output from the correlator 410.
  • the horizontal scale shows the correlation delay (in terms ofthe sequence samples).
  • the vertical scale on the left hand side (referred to as the confidence level cL) represents the value ofthe correlation peak normalized with respect to the standard deviation ofthe normally distributed correlation function.
  • the typical correlation is relatively flat with respect to cL, and centered about cL - 0.
  • the function contains two peaks, which are separated by pL (see equation 6) and extend upwards to cL values that are above the detection threshold when a watermark is present.
  • the correlation peaks are negative, the above statement applies to their absolute values.
  • the detection threshold value controls the false alarm rate.
  • detection criteria can be altered depending upon the desired use ofthe watermark signal and to take into account factors such as the original quality ofthe host signal and how badly the signal is likely to be corrupted during normal transmission.
  • the payload extractor unit 430 may subsequently be utilized to extract the payload (e.g. information content) from the detected watermark signal.
  • the payload e.g. information content
  • an estimate cL' ofthe circular shift cL (defined in equation (6)) is derived as the distance between the peaks .
  • the signs pi and 2 ofthe correlation peaks are determined, and hence r S jg n calculated from equation (7).
  • the present invention can be applied to add information to other types of signal, for instance information or multimedia signals, such as video and data signals.
  • the invention can be applied to watermarking schemes containing only one watermarking sequence (i.e. a 1-bit scheme), or to watermarking schemes containing multiple watermarking sequences. Such multiple sequences can be simultaneously or successively embedded within the host signal.

Landscapes

  • Engineering & Computer Science (AREA)
  • Signal Processing (AREA)
  • Physics & Mathematics (AREA)
  • Health & Medical Sciences (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Human Computer Interaction (AREA)
  • Computational Linguistics (AREA)
  • Acoustics & Sound (AREA)
  • Multimedia (AREA)
  • General Physics & Mathematics (AREA)
  • Editing Of Facsimile Originals (AREA)
  • Television Systems (AREA)
  • Image Processing (AREA)
  • Television Signal Processing For Recording (AREA)

Abstract

Method and apparatus are described for compensating for a linear time scale change in a received signal, so as to correctly rescale the frame sequence of the received signal. Firstly, an initial estimate of the sequence of symbols is extracted from the received signal. Successive estimates of correctly time scaled sequences of the symbols are then generated by interpolating the values of the initial estimates.

Description

Watermark time scale searching
The present invention relates to apparatus and methods for decoding information that has been embedded in information signals, such as audio, video or data signals.
Watermarking of information signals is a technique for the transmission of additional data along with the information signal. For instance, watermarking techniques can be used to embed copyright and copy control information into audio signals.
The main requirement of a watermarking scheme is that it is not observable (i.e. in the case of an audio signal, it is inaudible) whilst being robust to attacks to remove the watermark from the signal (e.g. removing the watermark will damage the signal). It will be appreciated that the robustness of a watermark will normally be a trade off against the quality ofthe signal in which the watermark is embedded. For instance, if a watermark is strongly embedded into an audio signal (and is thus difficult to remove) then it is likely that the quality ofthe audio signal will be reduced. In digital devices, it is typically assumed that there exists up to a 1% drift in sampling (clock) frequency. During transmission ofthe signal through an analog channel, this drift is normally manifested as a stretch or shrink in the time domain signal (i.e. a linear time scale change). A watermark embedded in the time domain (e.g. in an audio signal) will be affected by this time stretch or shrink as well, which can make watermark detection very difficult or even impossible. Thus, in the implementation of a robust watermarking scheme, it is extremely important to find solutions to such time scale modifications.
In known time domain watermarking schemes, any linear time scale change within the signal is resolved by repeatedly running the watermark detection (including repeating the extraction ofthe watermark from the host signal) for the different possible time scales, until all the possible time scales are exhausted, or detection is achieved. Performing such searches over the possible time scaling ranges requires a large computational overhead, and is thus costly in terms of both hardware and computational time. Consequently, real time implementation of a watermark detector utilizing such a time scale search technique is not feasible. In watermarking schemes implemented within the frequency domains, it is common to perform the scale search by modifying the frequency domain coefficients. For instance, this can be achieved by carefully shrinking or stretching the frequency domain samples. In principle, such a frequency domain solution could be directly applied to time domain watermark signals. However, since the watermarks are directly embedded in the time domain samples, the time scale search needs to be performed in the time domain as well. Normally, there are only a few thousand frequency domain samples, whilst the time domain signals contain samples in the order of millions. Consequently, such an application of the frequency domain solution to time domain signals is computationally too expensive.
It is an object ofthe present invention to provide a watermark decoding scheme for time domain watermarked signals that utilizes a time scale search that substantially addresses at least one ofthe problems ofthe prior art.
In a first aspect, the present invention provides a method of compensating for a linear time scale change in a received signal, the signal being modified by a sequence of symbols in the time domain, the method comprising the steps of: (a) extracting an initial estimate ofthe sequence of symbols from said received signal;(b) forming an estimate of a correctly time scaled sequence ofthe symbols by interpolating the values of said initial estimate. Preferably, step (b) is repeated so as to provide a range of estimates corresponding to different time scalings.
Preferably, said interpolation is at least one of zeroth order interpolation, linear interpolation, quadratic interpolation and cubic interpolation.
Preferably, the method further comprises the step of processing each estimate as though it were the correctly time scaled sequence ofthe symbols, so as to determine which estimate is the best estimate.
Preferably, the method further comprises the steps of correlating each of said estimates with a reference corresponding to said sequence of symbols; and taking the estimate with the maximum correlation peak as the best estimate. Preferably, said initial estimate ofthe sequence of symbols is stored in a buffer. Preferably, said buffer is of total length M, the total
Figure imgf000004_0001
bf siskin ..searghjs/
M conducted is Nη = — (77max - 77min ) where ηmin, ηmzx correspond respectively to the mimmum
and maximum likely time scale modifications ofthe signal.
Preferably, said initial estimates ofthe sequence of symbols comprises a sequence of Nb estimates for each symbol, each ofthe Nb estimates corresponding to a different time offset of a symbol.
Preferably, the scale search in the next detection window is adapted based on the information acquired during the current detection window.
Preferably, the scale space is searched using an optimal searching algorithm. Preferably, the searching algorithm is the grid refinement algorithm.
In another aspect, the present invention provides a computer program arranged to perform the method as described above.
In further aspects, the present invention provides a record carrier comprising the computer program, and a method of making available for downloading the computer program.
In another aspect, the present invention provides an apparatus arranged to compensate for a linear time scale change in a received signal, the signal being modified by a sequence of symbols in the time domain, the apparatus comprising: an extractor arranged to extract an initial estimate ofthe sequence of symbols from said received signal; and an interpolator arranged to form an estimate of a correctly time scaled sequence ofthe symbols by interpolating the values of said initial estimate.
Preferably, the apparatus further comprises a buffer arranged to store one or more of said estimates.
In another aspect, the present invention provides a decoder comprising the apparatus as described above.
For a better understanding ofthe invention, and to show how embodiments of the same may be carried into effect, reference will now be made, by way of example, to the accompanying diagrammatic drawings in which:
Figure 1 is a diagram illustrating a watermark embedding apparatus; Figure 2 shows a signal portion extraction filter H;
Figures 3 a and 3b show respectively the typical amplitude and phase responses as a function of frequency ofthe filter Η shown in Fig. 2; Figure 4 shows the payload embedding and watermark conditioning stage of the apparatus shown in Fig. 1 ;
Figure 5 is a diagram illustrating the details ofthe watermark conditioning apparatus He of Fig. 4, including charts ofthe associated signals at each stage; Figure 6a and 6b show two preferred alternative window shaping functions s(n) in the form of respectively a raised cosine function and a bi-phase function;
Figures 7a and 7b show respectively the frequency spectra for a watermark sequence conditioned with a raised cosine and a bi-phase shaping window function;
Figure 8 is a diagram illustrating a watermark detector in accordance with an embodiment of the present invention;
Figure 9 diagrammatically shows the whitening filter Hw of Fig. 8, for use in conjunction with a raised cosine shaping window function;
Figure 10 diagrammatically shows the whitening filter Hwof Fig. 8, for use in conjunction with a bi-phase window shaping function; Figure 11 shows details ofthe watermark symbol extraction and buffering processes in accordance with an embodiment ofthe present invention;
Figure 12 illustrates a sequence in which estimates of watermark symbols are collected from four buffers when there is no time scale modification;
Figures 13a and 13b illustrate the different sequences, according to an embodiment ofthe present invention, in which estimates of watermark symbols can be collected from four buffers when there is respectively a time stretch and a time shrink time scale modification;
Figure 14 shows an example of an efficient scale search technique based on the concept of grid refinement; and Figure 15 shows a typical shape ofthe correlation function output from the correlator ofthe watermark detector shown in Fig. 8.
Fig. 1 shows a block diagram ofthe apparatus required to perform the digital signal processing for embedding a multi-bit payload watermark w into a host signal x. A host signal x is provided at an input 12 ofthe apparatus. The host signal x is passed in the direction of output 14 via the adder 22. However, a replica ofthe host signal x (input 8) is split off in the direction ofthe multiplier 18, for carrying the watermark information. The watermark signal wc is obtained from the payload embedder and watermark conditioning apparatus 6, and derived from a reference finite length random sequence ws input to the payload embedder and watermark conditioning apparatus. The multiplier 18 is utilized to calculate the product ofthe watermark signal wc and the replica audio signal x. The resulting product, WcX is then passed via a gain controller 24 to the adder 22. The gain controller 24 is used to amplify or attenuate the signal by a gain factor α.
The gain factor α controls the trade off between the audibility and the robustness ofthe watermark. It may be a constant, or variable in at least one of time, frequency and space. The apparatus in Fig. 1 shows that, when α is variable, it can be automatically adapted via a signal analyzing unit 26 based upon the properties ofthe host signal x. Preferably, the gain α is automatically adapted, so as to minimize the impact on the signal quality, according to a properly chosen perceptibility cost- function, such as a psycho- acoustic model ofthe human auditory system (HAS) in case of an audio signal. Such a model is, for instance, described in the paper by E.Zwicker, "Audio Engineering and Psychoacoustics: Matching signals to the final receiver, the Human Auditory System", Journal ofthe Audio Engineering Society, Vol. 39, pp. Vol.115-126, March 1991.
In the following, an audio watermark is utilized, by way of example only, to describe this embodiment ofthe present invention.
The resulting watermark audio signal y is then obtained at the output 14 ofthe embedding apparatus 10 by adding an appropriately scaled version of the product of wc and x to the host signal: y[n] = x[n] + awc[n]x[n] . (1)
Preferably, the watermark wc is chosen such that when multiplied with x, it predominantly modifies the short time envelope of x. Fig. 2 shows one preferred embodiment in which the input 8 to the multiplier
18 in Fig. 1 is obtained by filtering a replica ofthe host signal x using a filter H in the filtering unit 15. If the filter output is denoted by xD, then according to this preferred embodiment, the watermark signal is generated by adding the product of xι and the watermark wc to the host signal x: y[n] = x + cxwc[n]xb[n] . (2) Let xb be defined such that xb = x - xb , and yb be defined such that y = yb + xb , then the envelope modulated portion yb ofthe watermarked signal y is given as yb[n] = (\ + wc[n])xb[n] (3)
Preferably, as shown in Fig. 3, the filter H is a linear phase band pass filter characterized by its lower cut-off frequency.//, and upper cut-off frequency fy. As can be seen in Fig. 3b, the filter H has a linear phase response with respect to frequency/ within the pass- band (BW). Thus, when H is a band pass filter, xb and xb are the in-band and out-of-band components ofthe host signal respectively. For optimum performance, it is preferable that the signals xb and xb are in phase. This is achieved by appropriately compensating for the phase distortion produced by filter H. In the case of a linear phase filter, the distortion is a simple time delay.
In Fig. 4, the details ofthe payload embedder and watermark conditioning unit 6 is shown. In this unit, the initial reference random sequence ws is converted into a multi-bit watermark signal wc. Firstly a finite length, preferably zero mean and uniformly distributed random sequence ws, from now on also referred to as the watermark seed signal, is generated using a random number generator with an initial seed S. As will be appreciated later, it is preferable that this initial seed S is known to both the embedder and the detector, such that a copy ofthe watermark signal can be generated at the detector for comparison purposes. This results in the sequence of length Lw ws[k] e [-1,1], for k=0,l,2 Lw-1 (4)
It should be noted that in some applications, the seed can be transmitted to the detector via an alternate channel or can be derived from the received signal using some predetermined protocol. Then the sequence ws is circularly shifted by the amounts dj and d using the circularly shifting unit 30 to obtain the random sequences WJJ and wd2 respectively. It will be appreciated that these two sequences (waj and vttø) are effectively a first sequence and a second sequence, with the second sequence being circularly shifted with respect to the first. Each sequence w*, i = 1 ,2, is subsequently multiplied with a respective sign bit rh in the multiplying unit 40, where η = +1 or -1. The respective values of ri and r2 remain constant, and only change when the payload ofthe watermark is changed. Each sequence is then converted into a periodic, slowly varying narrow-band signal wt of length LWTS by the watermark conditioning circuit 20 shown in Fig. 4. Finally, the slowly varying narrow-band signals wj and w2 are added with a relative delay Tr (where Tr<Ts) to give the multi-bit payload watermark signal wc. This is achieved by first delaying the signal w2 by the amount Tr using delaying unit 45 and subsequently by adding it to wj with the adding unit 50.
Fig. 5 shows the watermark conditioning apparatus 20 used in the payload embedder and watermark conditioning apparatus 6 in more detail. The watermark seed signal ws is input to the conditioning apparatus 20.
For convenience, the modification of only one ofthe sequences w^,- is shown in Fig. 5, but it will be appreciated that each ofthe sequences is modified in a similar manner, with the results being added to obtain the watermark signal wc.
As shown in Fig. 5, each watermark signal sequence Wdi[k], i=l,2 is applied to the input of a sample repeater 180. Chart 181 illustrates one ofthe sequences w^ as a sequence of values of random numbers between +1 and -ls with the sequence being of length Lw. The sample repeater repeats each value within the watermark seed signal sequence Ts times, so as to generate a rectangular pulse train signal. Ts is referred to as the watermark symbol period and represents the span ofthe watermark symbol in the audio signal. Chart 183 shows the results ofthe signal illustrated in chart 181 once it has passed through the sample repeater 180. A window shaping function s[n], such as a raised cosine window, is then applied to convert the rectangular pulse functions derived from w^ and vitø into slowly varying watermark sequence functions wjfnj and w2[n] respectively.
Chart 184 shows a typical raised cosine window shaping function, which is also of span Ts. The generated watermark sequences wj[n] and w2[n] are then added up with a relative delay Tr (where Tr<Ts) to give the multi-bit payload watermark signal wc[n] i.e.,
wc[n] = w,[n] + w2[n -Tr] (5)
The value of Tr is chosen such that the zero crossings of wt match the maximum amphtude points of w and vice-versa. Thus, for a raised cosine window shaping function Tr=TJ2, and for a bi-phase window shaping function Tr=T 4. For other window shaping functions, other values of Tr are possible. As will be appreciated by the below description, during detection the correlation of wc[n] will generate two correlation peaks that are separated by pL ' (as can be seen in Fig. 15). pL ' is an estimate ofthe circular shift pL between Wdi and wj2, which is part ofthe payload, and is defined as
Figure imgf000009_0001
In addition to pL, extra information can be encoded by changing the relative signs ofthe embedded watermarks.
In the detector, this is seen as a relative sign rsign between the correlation peaks. It may be defined as:
r,lg, - 2- + z + 3 S {0.1,2,3} (7)
where
Figure imgf000009_0002
are respectively estimates ofthe sign bits ri (input 80) and r2 (input 90) of Fig. 4, and cLi and cL2 are the values ofthe correlation peak corresponding to wj) and Wd2 respectively. The overall watermark payload pLw, for an error- free detection, is then given as a combination of rSjgn and pL:
pLw = (rsisn,pL) . (8)
The maximum information (Im<Y), in number of bits, that can be carried by a watermark sequence of length Lw is thus given by:
Figure imgf000009_0003
In such a scheme, the payload is immune to relative offset between the embedder and the detector, and also to possible time scale modifications.
The window shaping function has been identified as one ofthe main parameters that controls the robustness and audibility behavior ofthe present watermarking scheme. As illustrated in Figs. 6a and b, two examples of possible window shaping functions are herein described - a raised cosine function and a bi-phase function. It is preferable to use a bi-phase window function instead of a raised cosine window function, so as to obtain a quasi DC-free watermark signal. This is illustrated in Figs. 7a and 7b, showing the frequency spectra corresponding to a watermark sequence (in this case a sequence of afk] = {1,1,-1,1,-1,-1,}) conditioned with respectively a raised cosine and a bi-phase window shaping function. As can be seen, the frequency spectrum for the raised cosine conditioned watermark sequence has a maximum at frequency/ = 0, whilst the frequency spectrum for the bi-phase shaped watermark sequence has a minimum at/= 0 i.e. it has very little DC component.
Useful information is only contained in the non-DC component ofthe watermark. Consequently, for the same added watermark energy, a watermark conditioned with the bi-phase window will carry more useful information than one conditioned by the raised cosine window. As a result, the bi-phase window offers superior audibility performance for the same robustness or, conversely, it allows a better robustness for the same audibility quality. Such a bi-phase function could be utilized as a window shaping function for other watermarking schemes. In other words, a bi-phase function could be applied to reduce the DC component of signals (such as a watermark) that are to be incorporated into another signal.
Fig. 8 shows a block diagram of a watermark detector (200, 300, 400). The detector consists of three major stages: (a) the watermark symbol extraction stage (200), (b) the buffering and interpolation stage (300), and (c) the correlation and decision stage (400).
In the symbol extraction stage (200), the received watermarked signal y '[n] is processed to generate multiple (Nb) estimates ofthe watermarked sequence. These estimates ofthe watermark sequence are required to resolve time offset that may exist between the embedder and the detector, so that the watermark detector can synchronize to the watermark sequence inserted in the host signal.
In the buffering and interpolation stage (300), these estimates are demultiplexed into Nb separate buffers, and an interpolation is applied to each buffer to resolve time scale modifications that may have occurred, e.g. a drift in sampling (clock) frequency may have resulted in a stretch or shrink in the time domain signal (i.e. the watermark may have been stretched or shrunk).
In the correlation and decision stage (400), the content of each buffer is correlated with the reference watermark and the maximum correlation peaks are compared against a threshold to determine the likelihood of whether the watermark is indeed embedded within the received signal y'[n].
In order to maximize the accuracy ofthe watermark detection, the watermark detection process is typically carried out over a length of received signal y'fnj that is 3 to 4 times that ofthe watermark sequence length. Thus each watermark symbol to be detected can be constructed by taking the average of several estimates of said symbol. This averaging process is referred to as smoothing, and the number of times the averaging is done is referred to as the smoothing factor Sf. Let LD be the detection window length, defined as the length of the audio segment (in number of samples) over which a watermark detection truth-value is reported. Then, LD=SJLWTS, where Ts is the symbol period and Lw the number of symbols within the watermark sequence. During symbol extraction, a factor Ts decimation takes place in the energy computation stage. Thus, the length (Lb) of each buffer 320 within the buffering and interpolation stage is L =SfLw.
In the watermark symbol extraction stage 200 shown in Fig. 8, the incoming watermark signal y '[n] is input to the optional signal conditioning filter Hb(210). This filter 210 is typically a band pass filter and has the same behavior as the corresponding filter (H, 15) shown in Fig. 2. The output ofthe filter Hi, is y'bfnj and, assuming linearity within the transmission medium, it follows from equations (1) and (3):
Yb ["] « t[n] = (1 + a )xb[ (10)
Note that in the above expression, the possible time offset between the embedder and the detector is implicitly ignored. For ease of explanation ofthe general watermarking scheme principles, from now on, it is assumed that there is perfect synchronism between the embedder and the detector (i.e. no offset). Explanation is given however below in reference to Fig. 11 of how to compensate for time offset in accordance with the present invention.
Note that when no filter is used in the embedder (i.e., when Η=l) then Hb in the detector can also be omitted, or it can still be included to improve the detection performance. If Hb is omitted, then yb in equation (10) is replaced with y. The rest ofthe processing is the same.
We assume that the audio signal is divided into frames of length Ts, and that y'b,m[n] is the n-th sample ofthe m-th filtered frame signal. The energy Efm] corresponding to the m-th frame is thus:
Figure imgf000012_0001
Combining this with equation 10, it follows that:
+ ocwe[m])xb nf (12)
Figure imgf000012_0002
where we[m] is the m-th extracted watermark symbol and contains Nb time-multiplexed estimates ofthe embedded watermark sequences. Solving for we[m] in equation 12 and ignoring higher order terms of α, gives the following approximation:
Figure imgf000012_0003
In the watermark extraction stage 200 shown in Fig. 8, the output y ' [nJ ofthe filter Hb is provided as an input to a frame divider 220, which divides the audio signal into frames of length Ts i.e. into y 'b.mfnj, with the energy calculating unit 230 then being used to calculate the energy corresponding to each ofthe framed signals as per equation (12). The output of this energy calculation unit 230 is then provided as an input to the whitening stage Hw (240) which performs the function shown in equation 13 so as to provide an output we[m] . Alternative implementations (240A, 240B) of this whitening stage are illustrated in Figs. 9 and 10.
It will be realized that the denominator of equation 13 contains a term that requires knowledge ofthe host (original) signal x. As the signal x is not available to the detector, it means that in order to calculate we[m] then the denominator of equation 13 must be estimated. Below is described how such an estimation can be achieved for the two described window shaping functions (the raised cosine window shaping function and the biphase window shaping function), but it will equally be appreciated that the teaching could be extended to other window shaping functions.
In relation to the raised cosine window shaping function shown in Fig. 6(a), it has been realized that the audio envelope induced by the watermark contributes only to the noisy part ofthe energy function EfmJ. The slowly varying part (i.e. the low frequency component) is predominately due to the contribution ofthe envelope ofthe original audio signal x. Thus, equation 13 may be approximated by:
Figure imgf000013_0001
where "lowpassQ" is a low pass filter function. Thus, it will be appreciated that the whitening filter Hw for the raised cosine window shape in the function can be realized as shown in Fig. 9.
As can be seen, such a whitening filter Hw (240A) comprises an input 242A for receiving the signal EfmJ. A portion of this signal is then passed through the low pass filter 247 A to produce a low pass filtered energy signal EipfmJ, which in turn is provided as an input to the calculation stage 248 A along with the function EfmJ. The calculation stage 248 A then divides EfmJ by EipfmJ to calculate the extracted watermark symbol wefmj.
When a bi-phase window function is employed in the watermark conditioning stage ofthe embedder, a different approach should be utilized to estimate the envelope ofthe original audio, and hence to calculate we[mj. It will be seen by examination ofthe bi-phase window function shown in Fig.
6b, that when the audio envelope is modulated with such a window function, the first and the second halves ofthe frame are scaled in opposite directions. In the detector, this property is utilized to estimate the envelope energy of the host signal x.
Consequently, within the detector, each audio frame is first sub-divided into two halves. The energy functions corresponding to the first and second half-frames are hence given by
Figure imgf000013_0002
and
T.-i
£ ]= Σ f (16) n=T 2
respectively. As the envelope ofthe original audio is modulated in opposite directions within the two sub-frames, the original audio envelope can be approximated as the mean of EjfmJ and E2 m/. Further, the instantaneous modulation value can be taken as the difference between these two functions. Thus, for the bi-phase window function, the watermark wefmj can be approximated by:
Figure imgf000014_0001
Consequently, the whitening filter Hw (240B) in Fig. 8 for a bi-phase window shaping function can be realized as shown in Fig. 10. Inputs 242B and 243B respectively receive the energy functions ofthe first and second half frames EifmJ and E2 . Each energy function is then split up into two, and provided to adders 245B and 246B which respectively calculate EifmJ - E2[m], and EifmJ + EfmJ. Both of these calculated functions are then passed to the calculating unit 248B which divides the value from adder 245B by the value from 246B so as to calculate we[rø/, containing Nb time-multiplexed estimates ofthe embedded watermark sequences, in accordance with equation 17.
This output wefmj is then passed to the buffering and interpolation stage 300 (Fig. 8), where the signal is de-multiplexed by a de-multiplexer 310, buffered in buffers 320 of length Lb, so as to resolve a lack of synchronism between the embedder and the detector, and interpolated within the interpolation unit 330 so as to compensate for a time scale modification between the embedder and the detector.
In order to maximize the possible robustness of a watermark, it is important to make sure that the watermarking system is immune to both time offsets and drifts in sampling frequency between the embedder and the detector. In other words, the watermark detector must be able to synchronize to the watermark sequence inserted in the host signal. Fig. 11 illustrates the process carried out by the buffering and interpolation stage 300 to resolve the offset issue. The example described illustrates the process for resolving offset when a raised cosine window shaping function has been employed in the watermark embedding process. However, in principle the same technique is applicable when the bi-phase window shaping function has been used.
Referring to Fig. 11, after filtering by the filter Hb 210, the incoming audio signal stream y ' fnJ is separated into preferably overlapping frames 302 of effective length Ts by the frame divider 220. Preferably, to resolve possible offset between the embedder and the detector, each frame is divided into Nb sub-frames (304a, 304b,...,304x), and the above computations (equations (12) to (17)) are applied on a sub-frame basis.
Preferably, each sub-frame overlaps with an adjacent sub-frame. In the example shown, it can be seen that there is a 50% overlap (Ts/Nb) of each sub-frame (304a, 304b, ..., 304x), with each ofthe sub-frames being of length 2TS/Nb. When overlapping sub- frames are considered, the main frames are preferably longer than the symbol period Ts so as to allow inter-frame overlap as shown in Fig. 11.
The energy ofthe audio is then computed for each sub-frame by the whitening stage 240, and the resulting values are de-multiplexed into the Nb buffers 320 by the demultiplexer 310. Each one(i?Λ B2, .... Bm) ofthe buffers 320 will thus contain a sequence of values, with the first buffer Bj containing a sequence of values corresponding to the first sub- frame within each frame, the second buffer B2 containing a sequence of values corresponding to the second sub-frame within each frame etc. If wDi is the content of the i-th buffer, then it can be shown that: w D,[k] = we[k - Nb
Figure imgf000015_0001
e {0,...,Lb - 1} (18)
where Lb is the buffer length.
For a raised cosine window shaping function, the energy ofthe embedded watermark is concentrated near the center ofthe frame, such that the sub-frame best aligned with the center ofthe frame will result in a distinctly better estimate ofthe embedded watermark symbol than all the other sub-frames. Effectively, each buffer thus contains an estimate ofthe symbol sequence, the estimates corresponding to the sequences having different time offsets. The sub-frame best aligned with the center ofthe frame (i.e. the best estimate ofthe correctly aligned frame) is determined by correlating the contents of each buffer with the reference watermark sequence. The sequence with the maximum correlation peak value is chosen as the best estimate ofthe correctly aligned frame. The corresponding confidence level, as described below, is used to determine the truth-value ofthe detection. Preferably, the correlation process is halted once an estimated watermark sequence with a correlation peak above the defined threshold has been found. Typically, the length of each buffer is between 3 to 4 times the watermark sequence length Lw, and is thus typically of length between 2048 and 8192 symbols, and Na is typically within the range of 2 to 8.
The buffer is normally 3 to 4 times that ofthe watermark sequence so that each watermark symbol can be constructed by taking the averages of several estimates of said symbol. This averaging process is referred to as smoothing, and the number of times the averaging is done is referred to as the smoothing factor Sf. Thus, given the buffer length Lb and the watermark sequence length Lw, the smoothing factor Sf is such that:
Figure imgf000016_0001
In another preferred embodiment, the detector ref nes the parameters used in the offset search based upon the results of a previous search step. For instance, if a first series of estimates shows that the results stored in buffer B3 provide the best estimate ofthe information signal, then the next offset search (either on the same received signal, or on the signal received during the next detection window) is refined by shifting the position ofthe sub-frames towards the position of the best estimate sub-frame. The estimates of the sequence having zero offset can thus be iteratively improved.
As previously mentioned, there can exist a drift in sampling (clock) frequency in digital devices, which results in a stretch or shrink in the time domain signal.
For instance, consider an audio segment s of length L that is time scaled such that it's new length becomes L, = L(l+ 17) where η is the time scaling factor, with η being a constant such that 1+τj >0; for a time stretch η>0, and for a time shrink η<0.
When the signal is not time scale modified (η =0), Νb estimates ofthe watermark sequence are constructed by collecting the symbols stored in the Νb buffers separately. Fig. 12 illustrates four buffers (Bl, B2, B3, B4), each buffer shown as a row of boxes, with each box within a row indicating a separate location within the respective buffer. The sequences wπ, W12, wι3, w1 are respective estimates ofthe watermark sequence. In the example shown in Fig. 12, it is assumed that the signal is not time scale modified, and hence each estimate (wπ, w^, wI3, w] ) represents an estimate ofthe watermark sequence with different time offset.
Consequently, each estimate (that is passed to the correlator 410) is formed by sequentially collecting the entries from each buffer. For example, the first value in sequence π (wπ [1]) is collected from the first location of Bl, the second (wπ [2]) from the second location of Bl etc, with the final value (wπ [Lb]) being collected from the final location ofthe buffer. It will be appreciated that the arrows, which connect each box in a row to the neighboring box, show the direction in which values ofthe sequence estimates are collected from the buffer locations. It will also be appreciated that, whilst only eleven buffer locations are shown for each buffer, the size ofthe buffers in practice is likely to be significantly larger than this. For example, in the preferred embodiment, the length of each buffer is typically between 2048 and 8192 locations, with the number of buffers typically being between 2 and 8. However, in order to prevent overflow of buffers during time scale search, the actual buffer lengths are set to (l+|r)max|) times the typical lengths specified above, where ηmax is the expected maximum scaling factor.
When the received signal y ' [n] has been time scale modified, it is necessary to perform a time scale search in order to correctly estimate the watermark sequence. In the present invention, such a search is performed by systematically combining the extracted watermark sequence estimates (we[m]), preferably by systematically combining
(interpolating) the different estimates ofthe watermark sequences stored in the buffers.
Such time scale searches can be performed by utilizing any order of interpolation. In the following two preferred embodiments, two orders of interpolation will be described - the first order (linear) interpolation and the zero order interpolation. However, it will be appreciated that this technique can be extended to higher orders of interpolation e.g. quadratic and cubic interpolation.
In the first embodiment, estimates ofthe time scaled watermark sequence are provided by applying linear interpolation to the previously extracted estimates ofthe watermark sequence. To this end, it can be assumed that the intermediate values we[k] generated by the symbol extraction step shown in Fig. 8 are sequentially stored in a single buffer of length M in place ofthe Nb buffers. In other words, that the Nb buffers are multiplexed into a single buffer of length
Figure imgf000017_0001
where Lw and sr are as defined earlier. Let the so stretched sequence be represented by w/ It can now be assumed that wD represents discrete samples of an otherwise continuous function. During time scale modification, these discrete points are either pushed towards each other or stretched out. This in turn is translated to re-sampling of the watermark function. In this embodiment, re-sampling is realized via a linear interpolation technique. That is, given the watermark sequence
Figure imgf000018_0001
...,M, an interpolated watermark sequence wjfmj is generated as
Wj[m] = μwD( l(l + η)mj ) + (l -μ)wD( [(l +
Figure imgf000018_0002
) (20)
Where μ = f(l+ η)m 1- (1 + η)m, and [ 7 and bJaxe the floor and the ceiling operators, respectively. After the interpolation, the watermark sequences are folded back into the Nb buffers in a similar way to that shown in Fig. 11. Let the interpolated watermark sequence folded into the buffer b e {0,...,Nb-l} be denoted by w/bfkj, then it can be shown that ,M = rD( l(Nbk + b)(\ + η)i )+ (l -μ)wD{ [(Nbk + b)(l +
Figure imgf000018_0003
). (21)
Let for b=l, .... Nb, wo.bfkj be the pre-interpolation sequence stored in the ό-th buffer, and qpk e {], ...sjLw} and rpt e{l, ...NbJ be defined as
Figure imgf000018_0004
and
Figure imgf000018_0005
Then, it can be shown that wD ( (Nbk + b)(\ + η)j ) = wDιl}Λ [gbk ] . Putting this into equation (21), it follows that wItb[k] = μwDΛk [qbk[k]] + (1 - μ)wD,(rhk+i)[gbk + 1] (22)
Thus, the interpolated buffer entries can be calculated directly from the N& sequences wob, b=l,..,Nb (as shown in Fig. 8, being passed to the correlator 410), by solving equation (22). A further embodiment ofthe present invention will now be described, in which estimates ofthe time scaled watermark sequence are provided by applying zero order interpolation to the previously extracted estimates ofthe watermark sequence. This approach can be represented with equation (22) with μ = 1. In this case, the interpolation function can be written as
W/.*[*] = w. ttø*[*]] . (23)
where qpk e (J, ...sLw} and rpk efl, ...Nb} are as defined above.
A graphical interpretation of equation (23) is shown in Figs. 13a & b. Fig. 13a shows how the different estimates ofthe correct watermark sequences (wπ, w^, WI3, W1 ) are extracted from the buffers for a time stretch, whilst Fig. 13b shows similar information for a time shrink. As in Fig. 12, each row of boxes represents a respective buffer, with each box representing a location within each buffer. The arrows indicate the order in which the buffer contents are collected from the estimates ofthe watermark sequences. When the audio signal is time scale modified, the start and the end ofthe framing will gradually drift backward or forward, depending respectively upon whether the signal is time scale stretched or compressed. The watermark symbol combining stage according to this embodiment tracks the size ofthe drift. When the absolute value ofthe cumulative drift exceeds Ts Nb (where Nb is the number of buffers i.e. the number of consecutive symbols that represent a single watermark symbol), then the symbol collection sequence from the buffers is adjusted to provide the next best estimate ofthe symbol from the buffers. In other words, the buffer counters are incremented or decremented (depending on drift direction), and a circular rotation ofthe buffer pointer for each watermark sequence estimation (wn, wI2, w>/j, wu) is performed. Let k be the buffer entry counter, where £ is an integer representing each location within each buffer i.e. k=l represents the first location within each buffer, k=2 the second etc. If the estimates ofthe watermark sequence are being taken from the buffers with no time scale modification (as shown in Fig. 12), then it will be appreciated that the values in the first sequence can be represented by wnfkj. However, for time scaled estimates, assuming that an estimate η is being made
ofthe time scale, then when
Figure imgf000019_0001
the counter values and the buffers from which the watermark estimates are taken are changed.
If η is positive (time stretch), the counter for the first buffer is incremented. The ordering ofthe buffers is also circularly shifted (i.e. the watermark sequence estimate wu previously being taken from buffer one will now been taken from buffer four, the estimate from buffer two will now be taken from buffer one, the estimate from three will now be taken from buffer two, and the estimate from buffer four will now be taken from buffer three). A similar circular shift is also performed on the buffer counter k. This is shown diagrammatically in Fig. 13 a. If η is negative (time stretch), the counter for the first buffer is incremented, and the ordering ofthe buffers is circularly shifted (i.e. the watermark sequence estimate wn previously being taken from buffer one will now be taken from buffer two, the estimate from buffer two will now be taken from buffer three, the estimate from three will now be taken from buffer four, and the estimate from buffer four will now be taken from buffer one). A similar circular shift is also performed on the buffer counter k. This is shown diagrammatically in Fig. 13b.
After these circular shifts and adjustment to the buffer counters have been performed the symbol collection to form the different estimates ofthe watermark sequences continues from left to right until
Figure imgf000020_0001
« (n + l)/Nb (i.e. the next interchange position is reached). The process of buffer order interchanging and the sequential symbol collection is then repeated until the end ofthe buffer is reached.
Consequently, it will be appreciated that a zeroth order interpolation ofthe time scaled watermark sequence has been performed. In other words, the time scaled watermark sequence has been estimated by selecting those values from the original, non time scaled watermark sequence estimates that would most closely correspond to the temporal positions ofthe time scaled watermark sequence. By utilizing previously extracted estimates ofthe watermark sequence, such a technique efficiently resolves the problems of estimating correctly time scaled watermarks, with minimal cost in terms of computational overhead. Such estimates ofthe time scaled watermark sequence will then be passed to the correlator (410), so as to determine whether the predicted time shift η accurately represents the time shift ofthe received signal i.e. do the estimates provided to the correlator provide good correlation peaks. If not, then the time scale search will be repeated for a different estimated value i.e. a different value of η. Due to possible time scale modification, the detection truth-value (whether or not the signal includes a watermark) is determined only after the appropriate scale search has been conducted. Let Δη be the scale search step size and let us assume that we want the watermark to survive all the scale modifications in the interval [//"min, max]- The total number of visited scales is then given by T ^max ~ ?7min .. .. Aη (24)
To rninimize Nη it is preferred to find the maximum value of Δη that can still allow an exhaustive scale search. To this end, experimental results show that the detection performance is not significantly affected if the time scaling does not exceed half of the inverse ofthe buffer length. This means that, for an exhaustive scale search, Δη should be such that 2
Aη ≤
NbsfLw
Putting this into equation (24), it follows that it is preferable to conduct a search over
Figure imgf000021_0001
time scales in order to conduct an exhaustive scale search. Clearly, any scale search can be time consuming. Thus, the complexity issue and cost in computing overhead should be considered when choosing the watermark embedding parameters Nb, Sf and Lw.
In one preferred embodiment the scale search is adapted such that information acquired during detection is utilized to plan an optimum search in the subsequent detection windows. For example, the scale search in the next detection window is started around the current optimum scale.
An alternative embodiment illustrated in Fig. 14 provides a method for efficient walk through the scale space by grid refinement. The most straightforward solution is a linear search from the minimum scale towards the maximum scale by adding up an incremental step. Assuming correlation, and thus confidence level, does not change abruptly from one scale to the next, one can considerably reduce the amount of scales visited during the search by reducing the space granularity. As shown in Fig. 14, the algorithm starts at scale zero and is repeated until a minimum granularity is reached or the watermark is detected (i.e., a local maximum for the confidence level is found) and/or the confidence level exceeds a predetermined threshold. When one has an indication where to start the scale search (e.g. an initial estimation from a previous detection), a random or linear search around this scale may suffice.
As shown in Fig. 8, outputs (WDI, WD2, ... wDNb) from the buffering stage are passed to the interpolation stage and, after interpolation, the outputs (wn, wn, ... w/w,) of this stage, which are needed to resolve a possible time scale modification in the watermarked signal, are passed to the correlation and decision stage. All ofthe estimates (wn, wn,- wmb) ofthe watermark corresponding to the different possible offset values are passed to the correlation and decision stage 400.
The correlator 410 calculates the correlation of each estimate
Figure imgf000022_0001
...,Nb with respect to the reference watermark sequence wcfkj. Each respective correlation output corresponding to each estimate is then applied to the maximum detection unit 420 which determines which two estimates provided the maximum correlation peak values. These estimates are chosen as the ones that best fit the circularly shifted versions Wdi and Wd2 ofthe reference watermark. The correlation values for these estimated sequences are passed to the threshold detector and payload extractor unit 430. The reference watermark sequence ws used within the detector corresponds to
(a possibly circularly shifted version of) the original watermark sequence applied to the host signal. For instance, if the watermark signal was calculated using a random number generator with seed S within the embedder, then equally the detector can calculate the same random number sequence using the same random number generation algorithm and the same initial seed S so as to determine the watermark signal. Altematively, the watermark signal originally applied in the embedder and utilized by the detector as a reference could simply be any predetermined sequence.
Fig. 15 shows a typical shape of a correlation function as output from the correlator 410. The horizontal scale shows the correlation delay (in terms ofthe sequence samples). The vertical scale on the left hand side (referred to as the confidence level cL) represents the value ofthe correlation peak normalized with respect to the standard deviation ofthe normally distributed correlation function.
As can be seen, the typical correlation is relatively flat with respect to cL, and centered about cL - 0. However, the function contains two peaks, which are separated by pL (see equation 6) and extend upwards to cL values that are above the detection threshold when a watermark is present. When the correlation peaks are negative, the above statement applies to their absolute values.
A horizontal line (shown in the Fig. as being set at cL = 8.7) represents the detection threshold. The detection threshold value controls the false alarm rate.
Two kinds of false alarms exist: The false positive rate, defined as the probability of detecting a watermark in non watermarked items, and the false negative rate, which is defined as the probability of not detecting a watermark in watermarked items. Generally, the requirement ofthe false positive alarm is more stringent than that ofthe false negative. The scale on the right hand side of Fig. 11 illustrates the probability of a false positive alarm ?. As can be seen in the example shown, the probability of a false positive p=10' is equivalent to the threshold cL = 8.7, whilst/? = 10'83 is equivalent to cL = 20. After each detection interval, the detector determines whether the original watermark is present or whether it is not present, and on this basis outputs a "yes" or a "no" decision. If desired, to improve this decision making process, a number of detection windows may be considered. In such an instance, the false positive probability is a combination ofthe individual probabilities for each detection window considered, dependent upon the desired criteria. For instance, it could be determined that if the correlation function has two peaks above a threshold of cL = 7 on any two out of three detection intervals, then the watermark is deemed to be present. Such detection criteria can be altered depending upon the desired use ofthe watermark signal and to take into account factors such as the original quality ofthe host signal and how badly the signal is likely to be corrupted during normal transmission. The payload extractor unit 430 may subsequently be utilized to extract the payload (e.g. information content) from the detected watermark signal. Once the unit has estimated the two correlation peaks cLi and cL2 that exceed the detection threshold, an estimate cL' ofthe circular shift cL (defined in equation (6)) is derived as the distance between the peaks . Next, the signs pi and 2 ofthe correlation peaks are determined, and hence rSjgn calculated from equation (7). The overall watermark payload may then be calculated using equation (8). For instance, it can be seen in Fig. 15 that pL is the relative distance between the two peaks. Both peaks are positive i.e. pj = +1, and 2 = +1. From equation (7), rSjgn = 3. Consequently, the payload pLw = <3, pL>. It will be appreciated by the skilled person that various implementations not specifically described would be understood as falling within the scope ofthe present invention. For instance, whilst only the functionality ofthe detecting apparatus has been described, it will be appreciated that the apparatus could be realized as a digital circuit, an analog circuit, a computer program, or a combination thereof.
Equally, whilst the above embodiment has been described with reference to an audio signal, it will be appreciated that the present invention can be applied to add information to other types of signal, for instance information or multimedia signals, such as video and data signals. Further, it will be appreciated that the invention can be applied to watermarking schemes containing only one watermarking sequence (i.e. a 1-bit scheme), or to watermarking schemes containing multiple watermarking sequences. Such multiple sequences can be simultaneously or successively embedded within the host signal.
Within the specification it will be appreciated that the word "comprising" does not exclude other elements or steps, that "a" or "and" does not exclude a plurality, and that a single processor or other unit may fulfil the functions of several means recited in the claims.

Claims

CLAIMS:
1. A method of compensating for a linear time scale change in a received signal, the signal being modified by a sequence of symbols in the time domain, the method comprising the steps of:
(a) extracting an initial estimate ofthe sequence of symbols from said received signal;
(b) forming an estimate of a correctly time scaled sequence ofthe symbols by interpolating the values of said initial estimate.
2. A method as claimed in claim 1, wherein step (b) is repeated so as to provide a range of estimates corresponding to different time scalings.
3. A method as claimed in claim 1, wherein said interpolation is at least one of zeroth order interpolation, linear interpolation, quadratic interpolation and cubic interpolation.
4. A method as claimed in claim 1 , the method further comprising the step of processing each estimate as though it were the correctly time scaled sequence ofthe symbols, so as to determine which estimate is the best estimate.
5. A method as claimed in claim 1, the method further comprising the steps of correlating each of said estimates with a reference corresponding to said sequence of symbols; and taking the estimate with the maximum correlation peak as the best estimate.
6. A method as claimed in claim 1, wherein said initial estimate ofthe sequence of symbols is stored in a buffer.
7. A method as claimed in claim 6, wherein said buffer is of total length M, the
M total number of scale searches conducted is N7 = — (ηmaxmiτ>) where 77mιn, max correspond respectively to the minimum and maximum likely time scale modifications ofthe signal.
8. A method as claimed in claim 1, wherein said initial estimates ofthe sequence of symbols comprises a sequence of Nb estimates for each symbol, each ofthe Nb estimates corresponding to a different time offset of a symbol.
9. A method as claimed in claim 1, wherein the scale search in the next detection window is adapted based on the information acquired during the current detection window.
10. A method as claimed in claim 1, wherein the scale space is searched using an optimal searching algorithm.
11. A method as claimed in claim 10, wherein the searching algorithm is the grid refinement algorithm.
12. A computer program arranged to perform the method as claimed in claim 1.
13. A record carrier comprising the computer program as claimed in claim 12.
14. A method of making available for downloading a computer program as claimed in claim 12.
15. An apparatus arranged to compensate for a linear time scale change in a received signal, the signal being modified by a sequence of symbols in the time domain, the apparatus comprising: an extractor arranged to extract an initial estimate ofthe sequence of symbols from said received signal; and an interpolator arranged to form an estimate of a correctly time scaled sequence of the symbols by interpolating the values of said initial estimate.
16. An apparatus as claimed in claim 15, the apparatus further comprising a buffer arranged to store one or more of said estimates.
7. A decoder comprising the apparatus as claimed in claim 15.
PCT/IB2003/000794 2002-03-28 2003-02-26 Watermark time scale searching Ceased WO2003083859A2 (en)

Priority Applications (6)

Application Number Priority Date Filing Date Title
KR10-2004-7015275A KR20040097227A (en) 2002-03-28 2003-02-26 Watermark time scale searching
DE60308667T DE60308667T2 (en) 2002-03-28 2003-02-26 WATERMARK TIME SCALE SEARCH
JP2003581193A JP4302533B2 (en) 2002-03-28 2003-02-26 Search for watermark time scale
EP03710065A EP1493145B1 (en) 2002-03-28 2003-02-26 Watermark time scale searching
US10/509,411 US7266466B2 (en) 2002-03-28 2003-02-26 Watermark time scale searching
AU2003214489A AU2003214489A1 (en) 2002-03-28 2003-02-26 Watermark time scale searching

Applications Claiming Priority (2)

Application Number Priority Date Filing Date Title
EP02076202.7 2002-03-28
EP02076202 2002-03-28

Publications (2)

Publication Number Publication Date
WO2003083859A2 true WO2003083859A2 (en) 2003-10-09
WO2003083859A3 WO2003083859A3 (en) 2004-05-13

Family

ID=28459514

Family Applications (1)

Application Number Title Priority Date Filing Date
PCT/IB2003/000794 Ceased WO2003083859A2 (en) 2002-03-28 2003-02-26 Watermark time scale searching

Country Status (9)

Country Link
US (1) US7266466B2 (en)
EP (1) EP1493145B1 (en)
JP (1) JP4302533B2 (en)
KR (1) KR20040097227A (en)
CN (1) CN100354931C (en)
AT (1) ATE341072T1 (en)
AU (1) AU2003214489A1 (en)
DE (1) DE60308667T2 (en)
WO (1) WO2003083859A2 (en)

Cited By (3)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
EP1612771A1 (en) * 2004-06-29 2006-01-04 Koninklijke Philips Electronics N.V. Scale searching for watermark detection
US7266466B2 (en) 2002-03-28 2007-09-04 Koninklijke Philips Electronics N.V. Watermark time scale searching
US9305560B2 (en) 2010-04-26 2016-04-05 The Nielsen Company (Us), Llc Methods, apparatus and articles of manufacture to perform audio watermark decoding

Families Citing this family (13)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JP4727025B2 (en) * 2000-08-01 2011-07-20 オリンパス株式会社 Image display device
US20050220322A1 (en) * 2004-01-13 2005-10-06 Interdigital Technology Corporation Watermarks/signatures for wireless communications
US7912159B2 (en) * 2004-01-26 2011-03-22 Hewlett-Packard Development Company, L.P. Enhanced denoising system
EP1729285A1 (en) * 2005-06-02 2006-12-06 Deutsche Thomson-Brandt Gmbh Method and apparatus for watermarking an audio or video signal with watermark data using a spread spectrum
WO2008091697A1 (en) * 2007-01-25 2008-07-31 Arbitron, Inc. Research data gathering
US8170087B2 (en) * 2007-05-10 2012-05-01 Texas Instruments Incorporated Correlation coprocessor
US9466307B1 (en) * 2007-05-22 2016-10-11 Digimarc Corporation Robust spectral encoding and decoding methods
CN102144237B (en) * 2008-07-03 2014-10-22 美国唯美安视国际有限公司 Efficient watermarking approaches of compressed media
US9269363B2 (en) 2012-11-02 2016-02-23 Dolby Laboratories Licensing Corporation Audio data hiding based on perceptual masking and detection based on code multiplexing
EP2835799A1 (en) * 2013-08-08 2015-02-11 Thomson Licensing Method and apparatus for detecting a watermark symbol in a section of a received version of a watermarked audio signal
US9418395B1 (en) 2014-12-31 2016-08-16 The Nielsen Company (Us), Llc Power efficient detection of watermarks in media signals
WO2018208997A1 (en) 2017-05-09 2018-11-15 Verimatrix, Inc. Systems and methods of preparing multiple video streams for assembly with digital watermarking
US10347262B2 (en) 2017-10-18 2019-07-09 The Nielsen Company (Us), Llc Systems and methods to improve timestamp transition resolution

Family Cites Families (22)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US8505108B2 (en) * 1993-11-18 2013-08-06 Digimarc Corporation Authentication using a digital watermark
US20020009208A1 (en) * 1995-08-09 2002-01-24 Adnan Alattar Authentication of physical and electronic media objects using digital watermarks
US6516079B1 (en) * 2000-02-14 2003-02-04 Digimarc Corporation Digital watermark screening and detecting strategies
US5889868A (en) * 1996-07-02 1999-03-30 The Dice Company Optimization methods for the insertion, protection, and detection of digital watermarks in digitized data
US6684199B1 (en) * 1998-05-20 2004-01-27 Recording Industry Association Of America Method for minimizing pirating and/or unauthorized copying and/or unauthorized access of/to data on/from data media including compact discs and digital versatile discs, and system and data media for same
IL133236A0 (en) * 1999-11-30 2001-03-19 Ttr Technologies Ltd Copy-protected digital audio compact disc and method and system for producing same
US7305104B2 (en) * 2000-04-21 2007-12-04 Digimarc Corporation Authentication of identification documents using digital watermarks
WO2001099109A1 (en) * 2000-06-08 2001-12-27 Markany Inc. Watermark embedding and extracting method for protecting digital audio contents copyright and preventing duplication and apparatus using thereof
US6674876B1 (en) * 2000-09-14 2004-01-06 Digimarc Corporation Watermarking in the time-frequency domain
WO2002025662A1 (en) * 2000-09-20 2002-03-28 Koninklijke Philips Electronics N.V. Distribution of content
GB2377511B (en) * 2001-07-03 2005-05-11 Macrovision Europ Ltd The copy protection of digital data
CN100380493C (en) * 2001-09-05 2008-04-09 皇家飞利浦电子股份有限公司 Robust watermarking for direct stream digital signals
US6975745B2 (en) * 2001-10-25 2005-12-13 Digimarc Corporation Synchronizing watermark detectors in geometrically distorted signals
US7047187B2 (en) * 2002-02-27 2006-05-16 Matsushita Electric Industrial Co., Ltd. Method and apparatus for audio error concealment using data hiding
DE60320546T2 (en) * 2002-03-28 2008-11-13 Koninklijke Philips Electronics N.V. LABELING OF TIME RANGE WITH WATERMARK FOR MULTIMEDIA SIGNALS
CN100354931C (en) 2002-03-28 2007-12-12 皇家飞利浦电子股份有限公司 Watermark time scale searching
JP4290014B2 (en) * 2002-03-28 2009-07-01 コーニンクレッカ フィリップス エレクトロニクス エヌ ヴィ Decoding watermarked information signals
EP1493155A1 (en) * 2002-03-28 2005-01-05 Koninklijke Philips Electronics N.V. Window shaping functions for watermarking of multimedia signals
DE60326578D1 (en) * 2002-06-03 2009-04-23 Koninkl Philips Electronics Nv REINTERBATION OF WATERMARK IN MULTIMEDIA SIGNALS
US20050240767A1 (en) * 2002-06-03 2005-10-27 Lemma Aweke N Encoding and decoding of watermarks in independent channels
EP1552454B1 (en) * 2002-10-15 2014-07-23 Verance Corporation Media monitoring, management and information system
US20050165690A1 (en) * 2004-01-23 2005-07-28 Microsoft Corporation Watermarking via quantization of rational statistics of regions

Cited By (4)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US7266466B2 (en) 2002-03-28 2007-09-04 Koninklijke Philips Electronics N.V. Watermark time scale searching
EP1612771A1 (en) * 2004-06-29 2006-01-04 Koninklijke Philips Electronics N.V. Scale searching for watermark detection
WO2006003570A1 (en) * 2004-06-29 2006-01-12 Koninklijke Philips Electronics N.V. Scale searching for watermark detection
US9305560B2 (en) 2010-04-26 2016-04-05 The Nielsen Company (Us), Llc Methods, apparatus and articles of manufacture to perform audio watermark decoding

Also Published As

Publication number Publication date
CN100354931C (en) 2007-12-12
EP1493145A2 (en) 2005-01-05
DE60308667D1 (en) 2006-11-09
WO2003083859A3 (en) 2004-05-13
US20050177332A1 (en) 2005-08-11
DE60308667T2 (en) 2007-08-23
AU2003214489A1 (en) 2003-10-13
US7266466B2 (en) 2007-09-04
ATE341072T1 (en) 2006-10-15
KR20040097227A (en) 2004-11-17
JP2005522080A (en) 2005-07-21
JP4302533B2 (en) 2009-07-29
AU2003214489A8 (en) 2003-10-13
EP1493145B1 (en) 2006-09-27
CN1643574A (en) 2005-07-20

Similar Documents

Publication Publication Date Title
EP1493145B1 (en) Watermark time scale searching
JP3659321B2 (en) Digital watermarking method and system
EP1514268B1 (en) Re-embedding of watermarks in multimedia signals
EP1493154B1 (en) Time domain watermarking of multimedia signals
JP2004525430A (en) Digital watermark generation and detection
US20010032313A1 (en) Embedding a watermark in an information signal
EP1493155A1 (en) Window shaping functions for watermarking of multimedia signals
US20070036357A1 (en) Watermarking of multimedia signals
US7546466B2 (en) Decoding of watermarked information signals
JP2002305650A (en) Apparatus for detecting and recovering data
KR20030014329A (en) Watermarking
KR20030016381A (en) Watermarking

Legal Events

Date Code Title Description
AK Designated states

Kind code of ref document: A2

Designated state(s): AE AG AL AM AT AU AZ BA BB BG BR BY BZ CA CH CN CO CR CU CZ DE DK DM DZ EC EE ES FI GB GD GE GH GM HR HU ID IL IN IS JP KE KG KP KR KZ LC LK LR LS LT LU LV MA MD MG MK MN MW MX MZ NO NZ OM PH PL PT RO RU SC SD SE SG SK SL TJ TM TN TR TT TZ UA UG US UZ VC VN YU ZA ZM ZW

AL Designated countries for regional patents

Kind code of ref document: A2

Designated state(s): GH GM KE LS MW MZ SD SL SZ TZ UG ZM ZW AM AZ BY KG KZ MD RU TJ TM AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HU IE IT LU MC NL PT SE SI SK TR BF BJ CF CG CI CM GA GN GQ GW ML MR NE SN TD TG

121 Ep: the epo has been informed by wipo that ep was designated in this application
WWE Wipo information: entry into national phase

Ref document number: 2003710065

Country of ref document: EP

WWE Wipo information: entry into national phase

Ref document number: 10509411

Country of ref document: US

WWE Wipo information: entry into national phase

Ref document number: 1020047015275

Country of ref document: KR

WWE Wipo information: entry into national phase

Ref document number: 20038071614

Country of ref document: CN

Ref document number: 2003581193

Country of ref document: JP

WWP Wipo information: published in national office

Ref document number: 1020047015275

Country of ref document: KR

WWP Wipo information: published in national office

Ref document number: 2003710065

Country of ref document: EP

WWG Wipo information: grant in national office

Ref document number: 2003710065

Country of ref document: EP