Document (#28063)

Author
Batlle, E.
Neuschmied, H.
Uray, P.
Ackermann, G.
Title
Recognition and analysis of audio for copyright protection : the RAA project
Source
Journal of the American Society for Information Science and Technology. 55(2004) no.12, S.1084-1091
Year
2004
Abstract
Automatic generation of play lists for commercial broadcast radio stations has become a major research topic. Audio identification systems have been around for a while, and they show good performance for clean audio files. However, songs transmitted by commercial radio stations are highly distorted to cause greater impact an the casual listener. This impact helps increase the probability that the listener will stay tuned in, but the price we have to pay is a severe modification in the audio itself. This causes the failure of traditional identification systems. Another problem is the fact that songs are never played from the beginning to the end. Actually, they are put an the air several seconds after their real beginning and almost always under the voice of a speaker. The same thing happens at the end. In this article, we present the RAA project, which was conceived to deal with real broadcast audio problems. The idea behind this project is to extract automatically an audio fingerprint (the so-called AudioDNA) that identifies the fragment of audio. This AudioDNA has to be robust enough to appear almost the same under several degrees of distortion. Once this AudioDNA is extracted from the broadcast audio, a matching algorithm is able to find its fragments inside a database. With this approach, the system can find not only a whole song but also small fragments of it, even with high distortion caused by broadcast (and DJ) manipulations.
Footnote
Beitrag in einem Themenheft zur Musikerschließung und zum Musikretrieval
Field
Musik

Similar documents (author)

  1. Ackermann, A.: Zur Rolle der Inhaltsanalyse bei der Sacherschließung : theoretischer Anspruch und praktische Wirklichkeit in der RSWK (2001) 6.16
    6.157975 = sum of:
      6.157975 = weight(author_txt:ackermann in 3062) [ClassicSimilarity], result of:
        6.157975 = fieldWeight in 3062, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          9.85276 = idf(docFreq=5, maxDocs=41962)
          0.625 = fieldNorm(doc=3062)
    
  2. Ackermann, J.: Knuth-Morris-Pratt (2005) 6.16
    6.157975 = sum of:
      6.157975 = weight(author_txt:ackermann in 1991) [ClassicSimilarity], result of:
        6.157975 = fieldWeight in 1991, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          9.85276 = idf(docFreq=5, maxDocs=41962)
          0.625 = fieldNorm(doc=1991)
    
  3. Ackermann, U.; Schumann, N.: DissOnline Portal (2007) 4.93
    4.92638 = sum of:
      4.92638 = weight(author_txt:ackermann in 4405) [ClassicSimilarity], result of:
        4.92638 = fieldWeight in 4405, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          9.85276 = idf(docFreq=5, maxDocs=41962)
          0.5 = fieldNorm(doc=4405)
    
  4. Payome, T.; Ackermann-Stommel, K.: Berufen zum Teletutor? : Interview mit Kerstin Ackermann-Stommel (2005) 4.31
    4.3105826 = sum of:
      4.3105826 = weight(author_txt:ackermann in 4521) [ClassicSimilarity], result of:
        4.3105826 = fieldWeight in 4521, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          9.85276 = idf(docFreq=5, maxDocs=41962)
          0.4375 = fieldNorm(doc=4521)
    

Similar documents (content)

  1. Inskip, C.: Music information retrieval research (2011) 0.12
    0.119754724 = sum of:
      0.119754724 = product of:
        0.49897802 = sum of:
          0.016224682 = weight(abstract_txt:impact in 2014) [ClassicSimilarity], result of:
            0.016224682 = score(doc=2014,freq=1.0), product of:
              0.06396963 = queryWeight, product of:
                1.0215956 = boost
                4.6378245 = idf(docFreq=1103, maxDocs=41962)
                0.013501451 = queryNorm
              0.25363103 = fieldWeight in 2014, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6378245 = idf(docFreq=1103, maxDocs=41962)
                0.0546875 = fieldNorm(doc=2014)
          0.019420778 = weight(abstract_txt:find in 2014) [ClassicSimilarity], result of:
            0.019420778 = score(doc=2014,freq=1.0), product of:
              0.07211641 = queryWeight, product of:
                1.0846988 = boost
                4.9242997 = idf(docFreq=828, maxDocs=41962)
                0.013501451 = queryNorm
              0.26929763 = fieldWeight in 2014, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.9242997 = idf(docFreq=828, maxDocs=41962)
                0.0546875 = fieldNorm(doc=2014)
          0.012070635 = weight(abstract_txt:this in 2014) [ClassicSimilarity], result of:
            0.012070635 = score(doc=2014,freq=2.0), product of:
              0.06329302 = queryWeight, product of:
                1.901096 = boost
                2.4658763 = idf(docFreq=9687, maxDocs=41962)
                0.013501451 = queryNorm
              0.19071038 = fieldWeight in 2014, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                2.4658763 = idf(docFreq=9687, maxDocs=41962)
                0.0546875 = fieldNorm(doc=2014)
          0.11607296 = weight(abstract_txt:songs in 2014) [ClassicSimilarity], result of:
            0.11607296 = score(doc=2014,freq=1.0), product of:
              0.23750734 = queryWeight, product of:
                1.9684783 = boost
                8.936469 = idf(docFreq=14, maxDocs=41962)
                0.013501451 = queryNorm
              0.48871315 = fieldWeight in 2014, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.936469 = idf(docFreq=14, maxDocs=41962)
                0.0546875 = fieldNorm(doc=2014)
          0.14233075 = weight(abstract_txt:listener in 2014) [ClassicSimilarity], result of:
            0.14233075 = score(doc=2014,freq=1.0), product of:
              0.272096 = queryWeight, product of:
                2.106945 = boost
                9.565078 = idf(docFreq=7, maxDocs=41962)
                0.013501451 = queryNorm
              0.5230902 = fieldWeight in 2014, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.565078 = idf(docFreq=7, maxDocs=41962)
                0.0546875 = fieldNorm(doc=2014)
          0.19285823 = weight(abstract_txt:audio in 2014) [ClassicSimilarity], result of:
            0.19285823 = score(doc=2014,freq=1.0), product of:
              0.52889377 = queryWeight, product of:
                5.874979 = boost
                6.667786 = idf(docFreq=144, maxDocs=41962)
                0.013501451 = queryNorm
              0.36464456 = fieldWeight in 2014, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.667786 = idf(docFreq=144, maxDocs=41962)
                0.0546875 = fieldNorm(doc=2014)
        0.24 = coord(6/25)
    
  2. Turner, J.M.; Colinet, E.: Using audio description for indexing moving images (2004) 0.10
    0.095448375 = sum of:
      0.095448375 = product of:
        0.7954031 = sum of:
          0.012193184 = weight(abstract_txt:this in 4725) [ClassicSimilarity], result of:
            0.012193184 = score(doc=4725,freq=1.0), product of:
              0.06329302 = queryWeight, product of:
                1.901096 = boost
                2.4658763 = idf(docFreq=9687, maxDocs=41962)
                0.013501451 = queryNorm
              0.1926466 = fieldWeight in 4725, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                2.4658763 = idf(docFreq=9687, maxDocs=41962)
                0.078125 = fieldNorm(doc=4725)
          0.3060096 = weight(abstract_txt:broadcast in 4725) [ClassicSimilarity], result of:
            0.3060096 = score(doc=4725,freq=1.0), product of:
              0.45021683 = queryWeight, product of:
                3.832816 = boost
                8.700081 = idf(docFreq=18, maxDocs=41962)
                0.013501451 = queryNorm
              0.6796938 = fieldWeight in 4725, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.700081 = idf(docFreq=18, maxDocs=41962)
                0.078125 = fieldNorm(doc=4725)
          0.47720036 = weight(abstract_txt:audio in 4725) [ClassicSimilarity], result of:
            0.47720036 = score(doc=4725,freq=3.0), product of:
              0.52889377 = queryWeight, product of:
                5.874979 = boost
                6.667786 = idf(docFreq=144, maxDocs=41962)
                0.013501451 = queryNorm
              0.90226126 = fieldWeight in 4725, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                6.667786 = idf(docFreq=144, maxDocs=41962)
                0.078125 = fieldNorm(doc=4725)
        0.12 = coord(3/25)
    
  3. Hu, X.; Choi, K.; Downie, J.S.: ¬A framework for evaluating multimodal music mood classification (2017) 0.09
    0.09257803 = sum of:
      0.09257803 = product of:
        0.7714836 = sum of:
          0.025304226 = weight(abstract_txt:same in 355) [ClassicSimilarity], result of:
            0.025304226 = score(doc=355,freq=1.0), product of:
              0.06782405 = queryWeight, product of:
                1.051923 = boost
                4.775505 = idf(docFreq=961, maxDocs=41962)
                0.013501451 = queryNorm
              0.37308633 = fieldWeight in 355, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.775505 = idf(docFreq=961, maxDocs=41962)
                0.078125 = fieldNorm(doc=355)
          0.017243765 = weight(abstract_txt:this in 355) [ClassicSimilarity], result of:
            0.017243765 = score(doc=355,freq=2.0), product of:
              0.06329302 = queryWeight, product of:
                1.901096 = boost
                2.4658763 = idf(docFreq=9687, maxDocs=41962)
                0.013501451 = queryNorm
              0.2724434 = fieldWeight in 355, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                2.4658763 = idf(docFreq=9687, maxDocs=41962)
                0.078125 = fieldNorm(doc=355)
          0.7289356 = weight(abstract_txt:audio in 355) [ClassicSimilarity], result of:
            0.7289356 = score(doc=355,freq=7.0), product of:
              0.52889377 = queryWeight, product of:
                5.874979 = boost
                6.667786 = idf(docFreq=144, maxDocs=41962)
                0.013501451 = queryNorm
              1.3782269 = fieldWeight in 355, product of:
                2.6457512 = tf(freq=7.0), with freq of:
                  7.0 = termFreq=7.0
                6.667786 = idf(docFreq=144, maxDocs=41962)
                0.078125 = fieldNorm(doc=355)
        0.12 = coord(3/25)
    
  4. Huthwaite, A.: IASA Cataloguing Rules for Audiovisual Media with emphasis on sound recordings : project goal and progress report (1996) 0.09
    0.088410296 = sum of:
      0.088410296 = product of:
        0.7367525 = sum of:
          0.040875882 = weight(abstract_txt:project in 299) [ClassicSimilarity], result of:
            0.040875882 = score(doc=299,freq=1.0), product of:
              0.08541055 = queryWeight, product of:
                1.4457511 = boost
                4.3755994 = idf(docFreq=1434, maxDocs=41962)
                0.013501451 = queryNorm
              0.4785812 = fieldWeight in 299, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.3755994 = idf(docFreq=1434, maxDocs=41962)
                0.109375 = fieldNorm(doc=299)
          0.1503912 = weight(abstract_txt:radio in 299) [ClassicSimilarity], result of:
            0.1503912 = score(doc=299,freq=1.0), product of:
              0.17782165 = queryWeight, product of:
                1.7032737 = boost
                7.7324967 = idf(docFreq=49, maxDocs=41962)
                0.013501451 = queryNorm
              0.8457418 = fieldWeight in 299, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.7324967 = idf(docFreq=49, maxDocs=41962)
                0.109375 = fieldNorm(doc=299)
          0.54548544 = weight(abstract_txt:audio in 299) [ClassicSimilarity], result of:
            0.54548544 = score(doc=299,freq=2.0), product of:
              0.52889377 = queryWeight, product of:
                5.874979 = boost
                6.667786 = idf(docFreq=144, maxDocs=41962)
                0.013501451 = queryNorm
              1.0313705 = fieldWeight in 299, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                6.667786 = idf(docFreq=144, maxDocs=41962)
                0.109375 = fieldNorm(doc=299)
        0.12 = coord(3/25)
    
  5. Christel, M.G.: Automated metadata in multimedia information systems : creation, refinement, use in surrogates, and evaluation (2009) 0.08
    0.08016284 = sum of:
      0.08016284 = product of:
        0.50101775 = sum of:
          0.026046138 = weight(abstract_txt:under in 87) [ClassicSimilarity], result of:
            0.026046138 = score(doc=87,freq=1.0), product of:
              0.08023378 = queryWeight, product of:
                1.1441178 = boost
                5.1940494 = idf(docFreq=632, maxDocs=41962)
                0.013501451 = queryNorm
              0.32462808 = fieldWeight in 87, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.1940494 = idf(docFreq=632, maxDocs=41962)
                0.0625 = fieldNorm(doc=87)
          0.009754547 = weight(abstract_txt:this in 87) [ClassicSimilarity], result of:
            0.009754547 = score(doc=87,freq=1.0), product of:
              0.06329302 = queryWeight, product of:
                1.901096 = boost
                2.4658763 = idf(docFreq=9687, maxDocs=41962)
                0.013501451 = queryNorm
              0.15411727 = fieldWeight in 87, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                2.4658763 = idf(docFreq=9687, maxDocs=41962)
                0.0625 = fieldNorm(doc=87)
          0.24480768 = weight(abstract_txt:broadcast in 87) [ClassicSimilarity], result of:
            0.24480768 = score(doc=87,freq=1.0), product of:
              0.45021683 = queryWeight, product of:
                3.832816 = boost
                8.700081 = idf(docFreq=18, maxDocs=41962)
                0.013501451 = queryNorm
              0.54375505 = fieldWeight in 87, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.700081 = idf(docFreq=18, maxDocs=41962)
                0.0625 = fieldNorm(doc=87)
          0.22040941 = weight(abstract_txt:audio in 87) [ClassicSimilarity], result of:
            0.22040941 = score(doc=87,freq=1.0), product of:
              0.52889377 = queryWeight, product of:
                5.874979 = boost
                6.667786 = idf(docFreq=144, maxDocs=41962)
                0.013501451 = queryNorm
              0.41673663 = fieldWeight in 87, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.667786 = idf(docFreq=144, maxDocs=41962)
                0.0625 = fieldNorm(doc=87)
        0.16 = coord(4/25)