Document (#28060)

Author
Batlle, E.
Neuschmied, H.
Uray, P.
Ackermann, G.
Title
Recognition and analysis of audio for copyright protection : the RAA project
Source
Journal of the American Society for Information Science and Technology. 55(2004) no.12, S.1084-1091
Year
2004
Abstract
Automatic generation of play lists for commercial broadcast radio stations has become a major research topic. Audio identification systems have been around for a while, and they show good performance for clean audio files. However, songs transmitted by commercial radio stations are highly distorted to cause greater impact an the casual listener. This impact helps increase the probability that the listener will stay tuned in, but the price we have to pay is a severe modification in the audio itself. This causes the failure of traditional identification systems. Another problem is the fact that songs are never played from the beginning to the end. Actually, they are put an the air several seconds after their real beginning and almost always under the voice of a speaker. The same thing happens at the end. In this article, we present the RAA project, which was conceived to deal with real broadcast audio problems. The idea behind this project is to extract automatically an audio fingerprint (the so-called AudioDNA) that identifies the fragment of audio. This AudioDNA has to be robust enough to appear almost the same under several degrees of distortion. Once this AudioDNA is extracted from the broadcast audio, a matching algorithm is able to find its fragments inside a database. With this approach, the system can find not only a whole song but also small fragments of it, even with high distortion caused by broadcast (and DJ) manipulations.
Footnote
Beitrag in einem Themenheft zur Musikerschließung und zum Musikretrieval
Field
Musik

Similar documents (author)

  1. Ackermann, A.: Zur Rolle der Inhaltsanalyse bei der Sacherschließung : theoretischer Anspruch und praktische Wirklichkeit in der RSWK (2001) 6.18
    6.1812773 = sum of:
      6.1812773 = weight(author_txt:ackermann in 3059) [ClassicSimilarity], result of:
        6.1812773 = fieldWeight in 3059, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          9.890043 = idf(docFreq=5, maxDocs=43556)
          0.625 = fieldNorm(doc=3059)
    
  2. Ackermann, J.: Knuth-Morris-Pratt (2005) 6.18
    6.1812773 = sum of:
      6.1812773 = weight(author_txt:ackermann in 1988) [ClassicSimilarity], result of:
        6.1812773 = fieldWeight in 1988, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          9.890043 = idf(docFreq=5, maxDocs=43556)
          0.625 = fieldNorm(doc=1988)
    
  3. Ackermann, U.; Schumann, N.: DissOnline Portal (2007) 4.95
    4.9450216 = sum of:
      4.9450216 = weight(author_txt:ackermann in 4402) [ClassicSimilarity], result of:
        4.9450216 = fieldWeight in 4402, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          9.890043 = idf(docFreq=5, maxDocs=43556)
          0.5 = fieldNorm(doc=4402)
    
  4. Payome, T.; Ackermann-Stommel, K.: Berufen zum Teletutor? : Interview mit Kerstin Ackermann-Stommel (2005) 4.33
    4.326894 = sum of:
      4.326894 = weight(author_txt:ackermann in 4518) [ClassicSimilarity], result of:
        4.326894 = fieldWeight in 4518, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          9.890043 = idf(docFreq=5, maxDocs=43556)
          0.4375 = fieldNorm(doc=4518)
    

Similar documents (content)

  1. Inskip, C.: Music information retrieval research (2011) 0.12
    0.11925466 = sum of:
      0.11925466 = product of:
        0.49689445 = sum of:
          0.015894799 = weight(abstract_txt:impact in 2011) [ClassicSimilarity], result of:
            0.015894799 = score(doc=2011,freq=1.0), product of:
              0.063113645 = queryWeight, product of:
                1.0185289 = boost
                4.605149 = idf(docFreq=1183, maxDocs=43556)
                0.013455698 = queryNorm
              0.25184408 = fieldWeight in 2011, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.605149 = idf(docFreq=1183, maxDocs=43556)
                0.0546875 = fieldNorm(doc=2011)
          0.01914294 = weight(abstract_txt:find in 2011) [ClassicSimilarity], result of:
            0.01914294 = score(doc=2011,freq=1.0), product of:
              0.071442895 = queryWeight, product of:
                1.0836555 = boost
                4.8996105 = idf(docFreq=881, maxDocs=43556)
                0.013455698 = queryNorm
              0.26794744 = fieldWeight in 2011, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.8996105 = idf(docFreq=881, maxDocs=43556)
                0.0546875 = fieldNorm(doc=2011)
          0.011522733 = weight(abstract_txt:this in 2011) [ClassicSimilarity], result of:
            0.011522733 = score(doc=2011,freq=2.0), product of:
              0.061376616 = queryWeight, product of:
                1.8790884 = boost
                2.4274454 = idf(docFreq=10449, maxDocs=43556)
                0.013455698 = queryNorm
              0.18773815 = fieldWeight in 2011, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                2.4274454 = idf(docFreq=10449, maxDocs=43556)
                0.0546875 = fieldNorm(doc=2011)
          0.11761056 = weight(abstract_txt:songs in 2011) [ClassicSimilarity], result of:
            0.11761056 = score(doc=2011,freq=1.0), product of:
              0.2396537 = queryWeight, product of:
                1.9847407 = boost
                8.973753 = idf(docFreq=14, maxDocs=43556)
                0.013455698 = queryNorm
              0.4907521 = fieldWeight in 2011, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.973753 = idf(docFreq=14, maxDocs=43556)
                0.0546875 = fieldNorm(doc=2011)
          0.13886029 = weight(abstract_txt:listener in 2011) [ClassicSimilarity], result of:
            0.13886029 = score(doc=2011,freq=1.0), product of:
              0.26771456 = queryWeight, product of:
                2.0977209 = boost
                9.484578 = idf(docFreq=8, maxDocs=43556)
                0.013455698 = queryNorm
              0.51868784 = fieldWeight in 2011, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.484578 = idf(docFreq=8, maxDocs=43556)
                0.0546875 = fieldNorm(doc=2011)
          0.19386312 = weight(abstract_txt:audio in 2011) [ClassicSimilarity], result of:
            0.19386312 = score(doc=2011,freq=1.0), product of:
              0.53084785 = queryWeight, product of:
                5.907813 = boost
                6.6778564 = idf(docFreq=148, maxDocs=43556)
                0.013455698 = queryNorm
              0.36519527 = fieldWeight in 2011, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.6778564 = idf(docFreq=148, maxDocs=43556)
                0.0546875 = fieldNorm(doc=2011)
        0.24 = coord(6/25)
    
  2. Turner, J.M.; Colinet, E.: Using audio description for indexing moving images (2004) 0.10
    0.09552769 = sum of:
      0.09552769 = product of:
        0.79606414 = sum of:
          0.011639718 = weight(abstract_txt:this in 4722) [ClassicSimilarity], result of:
            0.011639718 = score(doc=4722,freq=1.0), product of:
              0.061376616 = queryWeight, product of:
                1.8790884 = boost
                2.4274454 = idf(docFreq=10449, maxDocs=43556)
                0.013455698 = queryNorm
              0.18964417 = fieldWeight in 4722, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                2.4274454 = idf(docFreq=10449, maxDocs=43556)
                0.078125 = fieldNorm(doc=4722)
          0.30473757 = weight(abstract_txt:broadcast in 4722) [ClassicSimilarity], result of:
            0.30473757 = score(doc=4722,freq=1.0), product of:
              0.44906852 = queryWeight, product of:
                3.842227 = boost
                8.68607 = idf(docFreq=19, maxDocs=43556)
                0.013455698 = queryNorm
              0.67859924 = fieldWeight in 4722, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.68607 = idf(docFreq=19, maxDocs=43556)
                0.078125 = fieldNorm(doc=4722)
          0.47968683 = weight(abstract_txt:audio in 4722) [ClassicSimilarity], result of:
            0.47968683 = score(doc=4722,freq=3.0), product of:
              0.53084785 = queryWeight, product of:
                5.907813 = boost
                6.6778564 = idf(docFreq=148, maxDocs=43556)
                0.013455698 = queryNorm
              0.90362394 = fieldWeight in 4722, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                6.6778564 = idf(docFreq=148, maxDocs=43556)
                0.078125 = fieldNorm(doc=4722)
        0.12 = coord(3/25)
    
  3. Hu, X.; Choi, K.; Downie, J.S.: ¬A framework for evaluating multimodal music mood classification (2017) 0.09
    0.09288675 = sum of:
      0.09288675 = product of:
        0.7740563 = sum of:
          0.02486153 = weight(abstract_txt:same in 352) [ClassicSimilarity], result of:
            0.02486153 = score(doc=352,freq=1.0), product of:
              0.06704564 = queryWeight, product of:
                1.0497768 = boost
                4.7464323 = idf(docFreq=1027, maxDocs=43556)
                0.013455698 = queryNorm
              0.37081504 = fieldWeight in 352, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.7464323 = idf(docFreq=1027, maxDocs=43556)
                0.078125 = fieldNorm(doc=352)
          0.016461046 = weight(abstract_txt:this in 352) [ClassicSimilarity], result of:
            0.016461046 = score(doc=352,freq=2.0), product of:
              0.061376616 = queryWeight, product of:
                1.8790884 = boost
                2.4274454 = idf(docFreq=10449, maxDocs=43556)
                0.013455698 = queryNorm
              0.26819736 = fieldWeight in 352, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                2.4274454 = idf(docFreq=10449, maxDocs=43556)
                0.078125 = fieldNorm(doc=352)
          0.7327337 = weight(abstract_txt:audio in 352) [ClassicSimilarity], result of:
            0.7327337 = score(doc=352,freq=7.0), product of:
              0.53084785 = queryWeight, product of:
                5.907813 = boost
                6.6778564 = idf(docFreq=148, maxDocs=43556)
                0.013455698 = queryNorm
              1.3803084 = fieldWeight in 352, product of:
                2.6457512 = tf(freq=7.0), with freq of:
                  7.0 = termFreq=7.0
                6.6778564 = idf(docFreq=148, maxDocs=43556)
                0.078125 = fieldNorm(doc=352)
        0.12 = coord(3/25)
    
  4. Huthwaite, A.: IASA Cataloguing Rules for Audiovisual Media with emphasis on sound recordings : project goal and progress report (1996) 0.09
    0.08892941 = sum of:
      0.08892941 = product of:
        0.74107844 = sum of:
          0.041235805 = weight(abstract_txt:project in 296) [ClassicSimilarity], result of:
            0.041235805 = score(doc=296,freq=1.0), product of:
              0.08593036 = queryWeight, product of:
                1.4555619 = boost
                4.3874254 = idf(docFreq=1471, maxDocs=43556)
                0.013455698 = queryNorm
              0.47987467 = fieldWeight in 296, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.3874254 = idf(docFreq=1471, maxDocs=43556)
                0.109375 = fieldNorm(doc=296)
          0.15151487 = weight(abstract_txt:radio in 296) [ClassicSimilarity], result of:
            0.15151487 = score(doc=296,freq=1.0), product of:
              0.17874618 = queryWeight, product of:
                1.7140762 = boost
                7.749977 = idf(docFreq=50, maxDocs=43556)
                0.013455698 = queryNorm
              0.84765375 = fieldWeight in 296, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.749977 = idf(docFreq=50, maxDocs=43556)
                0.109375 = fieldNorm(doc=296)
          0.54832774 = weight(abstract_txt:audio in 296) [ClassicSimilarity], result of:
            0.54832774 = score(doc=296,freq=2.0), product of:
              0.53084785 = queryWeight, product of:
                5.907813 = boost
                6.6778564 = idf(docFreq=148, maxDocs=43556)
                0.013455698 = queryNorm
              1.0329282 = fieldWeight in 296, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                6.6778564 = idf(docFreq=148, maxDocs=43556)
                0.109375 = fieldNorm(doc=296)
        0.12 = coord(3/25)
    
  5. Christel, M.G.: Automated metadata in multimedia information systems : creation, refinement, use in surrogates, and evaluation (2009) 0.08
    0.0800653 = sum of:
      0.0800653 = product of:
        0.5004082 = sum of:
          0.025748499 = weight(abstract_txt:under in 84) [ClassicSimilarity], result of:
            0.025748499 = score(doc=84,freq=1.0), product of:
              0.079639144 = queryWeight, product of:
                1.144129 = boost
                5.1730337 = idf(docFreq=670, maxDocs=43556)
                0.013455698 = queryNorm
              0.3233146 = fieldWeight in 84, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.1730337 = idf(docFreq=670, maxDocs=43556)
                0.0625 = fieldNorm(doc=84)
          0.009311774 = weight(abstract_txt:this in 84) [ClassicSimilarity], result of:
            0.009311774 = score(doc=84,freq=1.0), product of:
              0.061376616 = queryWeight, product of:
                1.8790884 = boost
                2.4274454 = idf(docFreq=10449, maxDocs=43556)
                0.013455698 = queryNorm
              0.15171534 = fieldWeight in 84, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                2.4274454 = idf(docFreq=10449, maxDocs=43556)
                0.0625 = fieldNorm(doc=84)
          0.24379005 = weight(abstract_txt:broadcast in 84) [ClassicSimilarity], result of:
            0.24379005 = score(doc=84,freq=1.0), product of:
              0.44906852 = queryWeight, product of:
                3.842227 = boost
                8.68607 = idf(docFreq=19, maxDocs=43556)
                0.013455698 = queryNorm
              0.5428794 = fieldWeight in 84, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.68607 = idf(docFreq=19, maxDocs=43556)
                0.0625 = fieldNorm(doc=84)
          0.22155786 = weight(abstract_txt:audio in 84) [ClassicSimilarity], result of:
            0.22155786 = score(doc=84,freq=1.0), product of:
              0.53084785 = queryWeight, product of:
                5.907813 = boost
                6.6778564 = idf(docFreq=148, maxDocs=43556)
                0.013455698 = queryNorm
              0.41736603 = fieldWeight in 84, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.6778564 = idf(docFreq=148, maxDocs=43556)
                0.0625 = fieldNorm(doc=84)
        0.16 = coord(4/25)