Document (#40225)

Author
Järvelin, A.
Keskustalo, H.
Sormunen, E.
Saastamoinen, M.
Kettunen, K.
Title
Information retrieval from historical newspaper collections in highly inflectional languages : a query expansion approach
Source
Journal of the Association for Information Science and Technology. 67(2016) no.12, S.2928-2946
Year
2016
Abstract
The aim of the study was to test whether query expansion by approximate string matching methods is beneficial in retrieval from historical newspaper collections in a language rich with compounds and inflectional forms (Finnish). First, approximate string matching methods were used to generate lists of index words most similar to contemporary query terms in a digitized newspaper collection from the 1800s. Top index word variants were categorized to estimate the appropriate query expansion ranges in the retrieval test. Second, the effectiveness of approximate string matching methods, automatically generated inflectional forms, and their combinations were measured in a Cranfield-style test. Finally, a detailed topic-level analysis of test results was conducted. In the index of historical newspaper collection the occurrences of a word typically spread to many linguistic and historical variants along with optical character recognition (OCR) errors. All query expansion methods improved the baseline results. Extensive expansion of around 30 variants for each query word was required to achieve the highest performance improvement. Query expansion based on approximate string matching was superior to using the inflectional forms of the query words, showing that coverage of the different types of variation is more important than precision in handling one type of variation.
Content
Vgl.: http://onlinelibrary.wiley.com/doi/10.1002/asi.23379/full.
Theme
Computerlinguistik
Semantisches Umfeld in Indexierung u. Retrieval
Form
Zeitungen

Similar documents (author)

  1. Järvelin, K.; Kristensen, J.; Niemi, T.; Sormunen, E.; Keskustalo, H.: ¬A deductive data model for query expansion (1996) 4.86
    4.856788 = sum of:
      4.856788 = sum of:
        1.1456299 = weight(author_txt:järvelin in 3695) [ClassicSimilarity], result of:
          1.1456299 = score(doc=3695,freq=1.0), product of:
            0.45612758 = queryWeight, product of:
              8.037259 = idf(docFreq=37, maxDocs=43254)
              0.056751635 = queryNorm
            2.5116434 = fieldWeight in 3695, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              8.037259 = idf(docFreq=37, maxDocs=43254)
              0.3125 = fieldNorm(doc=3695)
        1.7617167 = weight(author_txt:sormunen in 3695) [ClassicSimilarity], result of:
          1.7617167 = score(doc=3695,freq=1.0), product of:
            0.60768825 = queryWeight, product of:
              1.154243 = boost
              9.27695 = idf(docFreq=10, maxDocs=43254)
              0.056751635 = queryNorm
            2.899047 = fieldWeight in 3695, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              9.27695 = idf(docFreq=10, maxDocs=43254)
              0.3125 = fieldNorm(doc=3695)
        1.9494413 = weight(author_txt:keskustalo in 3695) [ClassicSimilarity], result of:
          1.9494413 = score(doc=3695,freq=1.0), product of:
            0.650125 = queryWeight, product of:
              1.1938652 = boost
              9.595404 = idf(docFreq=7, maxDocs=43254)
              0.056751635 = queryNorm
            2.9985638 = fieldWeight in 3695, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              9.595404 = idf(docFreq=7, maxDocs=43254)
              0.3125 = fieldNorm(doc=3695)
    
  2. Lehtokangas, R.; Keskustalo, H.; Järvelin, K.: Experiments with transitive dictionary translation and pseudo-relevance feedback using graded relevance assessments (2008) 2.48
    2.476057 = sum of:
      2.476057 = product of:
        3.7140853 = sum of:
          1.3747559 = weight(author_txt:järvelin in 3350) [ClassicSimilarity], result of:
            1.3747559 = score(doc=3350,freq=1.0), product of:
              0.45612758 = queryWeight, product of:
                8.037259 = idf(docFreq=37, maxDocs=43254)
                0.056751635 = queryNorm
              3.0139723 = fieldWeight in 3350, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.037259 = idf(docFreq=37, maxDocs=43254)
                0.375 = fieldNorm(doc=3350)
          2.3393295 = weight(author_txt:keskustalo in 3350) [ClassicSimilarity], result of:
            2.3393295 = score(doc=3350,freq=1.0), product of:
              0.650125 = queryWeight, product of:
                1.1938652 = boost
                9.595404 = idf(docFreq=7, maxDocs=43254)
                0.056751635 = queryNorm
              3.5982764 = fieldWeight in 3350, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.595404 = idf(docFreq=7, maxDocs=43254)
                0.375 = fieldNorm(doc=3350)
        0.6666667 = coord(2/3)
    
  3. Pirkola, A.; Hedlund, T.; Keskustalo, H.; Järvelin, K.: Dictionary-based cross-language information retrieval : problems, methods, and research findings (2001) 2.06
    2.063381 = sum of:
      2.063381 = product of:
        3.0950713 = sum of:
          1.1456299 = weight(author_txt:järvelin in 5909) [ClassicSimilarity], result of:
            1.1456299 = score(doc=5909,freq=1.0), product of:
              0.45612758 = queryWeight, product of:
                8.037259 = idf(docFreq=37, maxDocs=43254)
                0.056751635 = queryNorm
              2.5116434 = fieldWeight in 5909, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.037259 = idf(docFreq=37, maxDocs=43254)
                0.3125 = fieldNorm(doc=5909)
          1.9494413 = weight(author_txt:keskustalo in 5909) [ClassicSimilarity], result of:
            1.9494413 = score(doc=5909,freq=1.0), product of:
              0.650125 = queryWeight, product of:
                1.1938652 = boost
                9.595404 = idf(docFreq=7, maxDocs=43254)
                0.056751635 = queryNorm
              2.9985638 = fieldWeight in 5909, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.595404 = idf(docFreq=7, maxDocs=43254)
                0.3125 = fieldNorm(doc=5909)
        0.6666667 = coord(2/3)
    
  4. Toivonen, J.; Pirkola, A.; Keskustalo, H.; Visala, K.; Järvelin, K.: Translating cross-lingual spelling variants using transformation rules (2005) 2.06
    2.063381 = sum of:
      2.063381 = product of:
        3.0950713 = sum of:
          1.1456299 = weight(author_txt:järvelin in 3053) [ClassicSimilarity], result of:
            1.1456299 = score(doc=3053,freq=1.0), product of:
              0.45612758 = queryWeight, product of:
                8.037259 = idf(docFreq=37, maxDocs=43254)
                0.056751635 = queryNorm
              2.5116434 = fieldWeight in 3053, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.037259 = idf(docFreq=37, maxDocs=43254)
                0.3125 = fieldNorm(doc=3053)
          1.9494413 = weight(author_txt:keskustalo in 3053) [ClassicSimilarity], result of:
            1.9494413 = score(doc=3053,freq=1.0), product of:
              0.650125 = queryWeight, product of:
                1.1938652 = boost
                9.595404 = idf(docFreq=7, maxDocs=43254)
                0.056751635 = queryNorm
              2.9985638 = fieldWeight in 3053, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.595404 = idf(docFreq=7, maxDocs=43254)
                0.3125 = fieldNorm(doc=3053)
        0.6666667 = coord(2/3)
    
  5. Ferro, N.; Silvello, G.; Keskustalo, H.; Pirkola, A.; Järvelin, K.: ¬The twist measure for IR evaluation : taking user's effort into account (2016) 2.06
    2.063381 = sum of:
      2.063381 = product of:
        3.0950713 = sum of:
          1.1456299 = weight(author_txt:järvelin in 4236) [ClassicSimilarity], result of:
            1.1456299 = score(doc=4236,freq=1.0), product of:
              0.45612758 = queryWeight, product of:
                8.037259 = idf(docFreq=37, maxDocs=43254)
                0.056751635 = queryNorm
              2.5116434 = fieldWeight in 4236, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.037259 = idf(docFreq=37, maxDocs=43254)
                0.3125 = fieldNorm(doc=4236)
          1.9494413 = weight(author_txt:keskustalo in 4236) [ClassicSimilarity], result of:
            1.9494413 = score(doc=4236,freq=1.0), product of:
              0.650125 = queryWeight, product of:
                1.1938652 = boost
                9.595404 = idf(docFreq=7, maxDocs=43254)
                0.056751635 = queryNorm
              2.9985638 = fieldWeight in 4236, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.595404 = idf(docFreq=7, maxDocs=43254)
                0.3125 = fieldNorm(doc=4236)
        0.6666667 = coord(2/3)
    

Similar documents (content)

  1. French, J.C.; Powell, A.L.; Schulman, E.: Using clustering strategies for creating authority files (2000) 0.33
    0.33382362 = sum of:
      0.33382362 = product of:
        1.0431988 = sum of:
          0.008284088 = weight(abstract_txt:from in 6812) [ClassicSimilarity], result of:
            0.008284088 = score(doc=6812,freq=1.0), product of:
              0.038134523 = queryWeight, product of:
                1.1144351 = boost
                2.7805862 = idf(docFreq=7289, maxDocs=43254)
                0.012306292 = queryNorm
              0.2172333 = fieldWeight in 6812, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                2.7805862 = idf(docFreq=7289, maxDocs=43254)
                0.078125 = fieldNorm(doc=6812)
          0.016098538 = weight(abstract_txt:retrieval in 6812) [ClassicSimilarity], result of:
            0.016098538 = score(doc=6812,freq=1.0), product of:
              0.05938537 = queryWeight, product of:
                1.3907061 = boost
                3.4699 = idf(docFreq=3658, maxDocs=43254)
                0.012306292 = queryNorm
              0.27108592 = fieldWeight in 6812, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.4699 = idf(docFreq=3658, maxDocs=43254)
                0.078125 = fieldNorm(doc=6812)
          0.06204923 = weight(abstract_txt:word in 6812) [ClassicSimilarity], result of:
            0.06204923 = score(doc=6812,freq=1.0), product of:
              0.14598653 = queryWeight, product of:
                2.1804795 = boost
                5.4404345 = idf(docFreq=509, maxDocs=43254)
                0.012306292 = queryNorm
              0.42503393 = fieldWeight in 6812, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.4404345 = idf(docFreq=509, maxDocs=43254)
                0.078125 = fieldNorm(doc=6812)
          0.08980125 = weight(abstract_txt:forms in 6812) [ClassicSimilarity], result of:
            0.08980125 = score(doc=6812,freq=2.0), product of:
              0.14825185 = queryWeight, product of:
                2.197332 = boost
                5.4824824 = idf(docFreq=488, maxDocs=43254)
                0.012306292 = queryNorm
              0.60573447 = fieldWeight in 6812, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.4824824 = idf(docFreq=488, maxDocs=43254)
                0.078125 = fieldNorm(doc=6812)
          0.1625919 = weight(abstract_txt:variants in 6812) [ClassicSimilarity], result of:
            0.1625919 = score(doc=6812,freq=1.0), product of:
              0.27747324 = queryWeight, product of:
                3.006119 = boost
                7.500458 = idf(docFreq=64, maxDocs=43254)
                0.012306292 = queryNorm
              0.58597326 = fieldWeight in 6812, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.500458 = idf(docFreq=64, maxDocs=43254)
                0.078125 = fieldNorm(doc=6812)
          0.16154483 = weight(abstract_txt:matching in 6812) [ClassicSimilarity], result of:
            0.16154483 = score(doc=6812,freq=2.0), product of:
              0.24135342 = queryWeight, product of:
                3.2373652 = boost
                6.058074 = idf(docFreq=274, maxDocs=43254)
                0.012306292 = queryNorm
              0.6693289 = fieldWeight in 6812, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                6.058074 = idf(docFreq=274, maxDocs=43254)
                0.078125 = fieldNorm(doc=6812)
          0.18890283 = weight(abstract_txt:string in 6812) [ClassicSimilarity], result of:
            0.18890283 = score(doc=6812,freq=1.0), product of:
              0.3375155 = queryWeight, product of:
                3.8283517 = boost
                7.1639853 = idf(docFreq=90, maxDocs=43254)
                0.012306292 = queryNorm
              0.55968636 = fieldWeight in 6812, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.1639853 = idf(docFreq=90, maxDocs=43254)
                0.078125 = fieldNorm(doc=6812)
          0.35392615 = weight(abstract_txt:approximate in 6812) [ClassicSimilarity], result of:
            0.35392615 = score(doc=6812,freq=2.0), product of:
              0.4071301 = queryWeight, product of:
                4.2046666 = boost
                7.8681827 = idf(docFreq=44, maxDocs=43254)
                0.012306292 = queryNorm
              0.86931956 = fieldWeight in 6812, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                7.8681827 = idf(docFreq=44, maxDocs=43254)
                0.078125 = fieldNorm(doc=6812)
        0.32 = coord(8/25)
    
  2. Galvez, C.; Moya-Anegón, F.: Approximate personal name-matching through finite-state graphs (2007) 0.33
    0.33023912 = sum of:
      0.33023912 = product of:
        0.91733086 = sum of:
          0.011478769 = weight(abstract_txt:from in 2615) [ClassicSimilarity], result of:
            0.011478769 = score(doc=2615,freq=3.0), product of:
              0.038134523 = queryWeight, product of:
                1.1144351 = boost
                2.7805862 = idf(docFreq=7289, maxDocs=43254)
                0.012306292 = queryNorm
              0.30100727 = fieldWeight in 2615, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                2.7805862 = idf(docFreq=7289, maxDocs=43254)
                0.0625 = fieldNorm(doc=2615)
          0.012878831 = weight(abstract_txt:retrieval in 2615) [ClassicSimilarity], result of:
            0.012878831 = score(doc=2615,freq=1.0), product of:
              0.05938537 = queryWeight, product of:
                1.3907061 = boost
                3.4699 = idf(docFreq=3658, maxDocs=43254)
                0.012306292 = queryNorm
              0.21686874 = fieldWeight in 2615, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.4699 = idf(docFreq=3658, maxDocs=43254)
                0.0625 = fieldNorm(doc=2615)
          0.046498016 = weight(abstract_txt:index in 2615) [ClassicSimilarity], result of:
            0.046498016 = score(doc=2615,freq=2.0), product of:
              0.11092808 = queryWeight, product of:
                1.9007121 = boost
                4.7423973 = idf(docFreq=1024, maxDocs=43254)
                0.012306292 = queryNorm
              0.41917264 = fieldWeight in 2615, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.7423973 = idf(docFreq=1024, maxDocs=43254)
                0.0625 = fieldNorm(doc=2615)
          0.101598516 = weight(abstract_txt:forms in 2615) [ClassicSimilarity], result of:
            0.101598516 = score(doc=2615,freq=4.0), product of:
              0.14825185 = queryWeight, product of:
                2.197332 = boost
                5.4824824 = idf(docFreq=488, maxDocs=43254)
                0.012306292 = queryNorm
              0.6853103 = fieldWeight in 2615, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                5.4824824 = idf(docFreq=488, maxDocs=43254)
                0.0625 = fieldNorm(doc=2615)
          0.042013012 = weight(abstract_txt:methods in 2615) [ClassicSimilarity], result of:
            0.042013012 = score(doc=2615,freq=2.0), product of:
              0.114109196 = queryWeight, product of:
                2.2260008 = boost
                4.1655097 = idf(docFreq=1824, maxDocs=43254)
                0.012306292 = queryNorm
              0.3681825 = fieldWeight in 2615, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.1655097 = idf(docFreq=1824, maxDocs=43254)
                0.0625 = fieldNorm(doc=2615)
          0.26014704 = weight(abstract_txt:variants in 2615) [ClassicSimilarity], result of:
            0.26014704 = score(doc=2615,freq=4.0), product of:
              0.27747324 = queryWeight, product of:
                3.006119 = boost
                7.500458 = idf(docFreq=64, maxDocs=43254)
                0.012306292 = queryNorm
              0.9375572 = fieldWeight in 2615, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                7.500458 = idf(docFreq=64, maxDocs=43254)
                0.0625 = fieldNorm(doc=2615)
          0.091383554 = weight(abstract_txt:matching in 2615) [ClassicSimilarity], result of:
            0.091383554 = score(doc=2615,freq=1.0), product of:
              0.24135342 = queryWeight, product of:
                3.2373652 = boost
                6.058074 = idf(docFreq=274, maxDocs=43254)
                0.012306292 = queryNorm
              0.37862962 = fieldWeight in 2615, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.058074 = idf(docFreq=274, maxDocs=43254)
                0.0625 = fieldNorm(doc=2615)
          0.15112226 = weight(abstract_txt:string in 2615) [ClassicSimilarity], result of:
            0.15112226 = score(doc=2615,freq=1.0), product of:
              0.3375155 = queryWeight, product of:
                3.8283517 = boost
                7.1639853 = idf(docFreq=90, maxDocs=43254)
                0.012306292 = queryNorm
              0.44774908 = fieldWeight in 2615, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.1639853 = idf(docFreq=90, maxDocs=43254)
                0.0625 = fieldNorm(doc=2615)
          0.20021087 = weight(abstract_txt:approximate in 2615) [ClassicSimilarity], result of:
            0.20021087 = score(doc=2615,freq=1.0), product of:
              0.4071301 = queryWeight, product of:
                4.2046666 = boost
                7.8681827 = idf(docFreq=44, maxDocs=43254)
                0.012306292 = queryNorm
              0.49176142 = fieldWeight in 2615, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.8681827 = idf(docFreq=44, maxDocs=43254)
                0.0625 = fieldNorm(doc=2615)
        0.36 = coord(9/25)
    
  3. Pirkola, A.; Puolamäki, D.; Järvelin, K.: Applying query structuring in cross-language retrieval (2003) 0.30
    0.30280733 = sum of:
      0.30280733 = product of:
        0.9462729 = sum of:
          0.10572249 = weight(abstract_txt:finnish in 3075) [ClassicSimilarity], result of:
            0.10572249 = score(doc=3075,freq=3.0), product of:
              0.10011963 = queryWeight, product of:
                1.0425445 = boost
                7.803644 = idf(docFreq=47, maxDocs=43254)
                0.012306292 = queryNorm
              1.0559616 = fieldWeight in 3075, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                7.803644 = idf(docFreq=47, maxDocs=43254)
                0.078125 = fieldNorm(doc=3075)
          0.027883485 = weight(abstract_txt:retrieval in 3075) [ClassicSimilarity], result of:
            0.027883485 = score(doc=3075,freq=3.0), product of:
              0.05938537 = queryWeight, product of:
                1.3907061 = boost
                3.4699 = idf(docFreq=3658, maxDocs=43254)
                0.012306292 = queryNorm
              0.46953458 = fieldWeight in 3075, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                3.4699 = idf(docFreq=3658, maxDocs=43254)
                0.078125 = fieldNorm(doc=3075)
          0.038498204 = weight(abstract_txt:were in 3075) [ClassicSimilarity], result of:
            0.038498204 = score(doc=3075,freq=4.0), product of:
              0.06690042 = queryWeight, product of:
                1.4760804 = boost
                3.6829145 = idf(docFreq=2956, maxDocs=43254)
                0.012306292 = queryNorm
              0.57545537 = fieldWeight in 3075, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                3.6829145 = idf(docFreq=2956, maxDocs=43254)
                0.078125 = fieldNorm(doc=3075)
          0.09399011 = weight(abstract_txt:test in 3075) [ClassicSimilarity], result of:
            0.09399011 = score(doc=3075,freq=2.0), product of:
              0.16820782 = queryWeight, product of:
                2.702639 = boost
                5.057442 = idf(docFreq=747, maxDocs=43254)
                0.012306292 = queryNorm
              0.5587737 = fieldWeight in 3075, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.057442 = idf(docFreq=747, maxDocs=43254)
                0.078125 = fieldNorm(doc=3075)
          0.1625919 = weight(abstract_txt:variants in 3075) [ClassicSimilarity], result of:
            0.1625919 = score(doc=3075,freq=1.0), product of:
              0.27747324 = queryWeight, product of:
                3.006119 = boost
                7.500458 = idf(docFreq=64, maxDocs=43254)
                0.012306292 = queryNorm
              0.58597326 = fieldWeight in 3075, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.500458 = idf(docFreq=64, maxDocs=43254)
                0.078125 = fieldNorm(doc=3075)
          0.11422945 = weight(abstract_txt:matching in 3075) [ClassicSimilarity], result of:
            0.11422945 = score(doc=3075,freq=1.0), product of:
              0.24135342 = queryWeight, product of:
                3.2373652 = boost
                6.058074 = idf(docFreq=274, maxDocs=43254)
                0.012306292 = queryNorm
              0.47328705 = fieldWeight in 3075, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.058074 = idf(docFreq=274, maxDocs=43254)
                0.078125 = fieldNorm(doc=3075)
          0.18470314 = weight(abstract_txt:newspaper in 3075) [ClassicSimilarity], result of:
            0.18470314 = score(doc=3075,freq=1.0), product of:
              0.33249435 = queryWeight, product of:
                3.7997682 = boost
                7.110497 = idf(docFreq=95, maxDocs=43254)
                0.012306292 = queryNorm
              0.5555076 = fieldWeight in 3075, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.110497 = idf(docFreq=95, maxDocs=43254)
                0.078125 = fieldNorm(doc=3075)
          0.2186541 = weight(abstract_txt:query in 3075) [ClassicSimilarity], result of:
            0.2186541 = score(doc=3075,freq=4.0), product of:
              0.29532248 = queryWeight, product of:
                5.0644026 = boost
                4.738502 = idf(docFreq=1028, maxDocs=43254)
                0.012306292 = queryNorm
              0.74039096 = fieldWeight in 3075, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                4.738502 = idf(docFreq=1028, maxDocs=43254)
                0.078125 = fieldNorm(doc=3075)
        0.32 = coord(8/25)
    
  4. Bellaachia, A.; Amor-Tijani, G.: Proper nouns in English-Arabic cross language information retrieval (2008) 0.29
    0.28714204 = sum of:
      0.28714204 = product of:
        0.8973189 = sum of:
          0.012878831 = weight(abstract_txt:retrieval in 4373) [ClassicSimilarity], result of:
            0.012878831 = score(doc=4373,freq=1.0), product of:
              0.05938537 = queryWeight, product of:
                1.3907061 = boost
                3.4699 = idf(docFreq=3658, maxDocs=43254)
                0.012306292 = queryNorm
              0.21686874 = fieldWeight in 4373, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.4699 = idf(docFreq=3658, maxDocs=43254)
                0.0625 = fieldNorm(doc=4373)
          0.054632217 = weight(abstract_txt:words in 4373) [ClassicSimilarity], result of:
            0.054632217 = score(doc=4373,freq=3.0), product of:
              0.094259165 = queryWeight, product of:
                1.4305787 = boost
                5.354077 = idf(docFreq=555, maxDocs=43254)
                0.012306292 = queryNorm
              0.5795958 = fieldWeight in 4373, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                5.354077 = idf(docFreq=555, maxDocs=43254)
                0.0625 = fieldNorm(doc=4373)
          0.032879066 = weight(abstract_txt:index in 4373) [ClassicSimilarity], result of:
            0.032879066 = score(doc=4373,freq=1.0), product of:
              0.11092808 = queryWeight, product of:
                1.9007121 = boost
                4.7423973 = idf(docFreq=1024, maxDocs=43254)
                0.012306292 = queryNorm
              0.29639983 = fieldWeight in 4373, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.7423973 = idf(docFreq=1024, maxDocs=43254)
                0.0625 = fieldNorm(doc=4373)
          0.13007352 = weight(abstract_txt:variants in 4373) [ClassicSimilarity], result of:
            0.13007352 = score(doc=4373,freq=1.0), product of:
              0.27747324 = queryWeight, product of:
                3.006119 = boost
                7.500458 = idf(docFreq=64, maxDocs=43254)
                0.012306292 = queryNorm
              0.4687786 = fieldWeight in 4373, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.500458 = idf(docFreq=64, maxDocs=43254)
                0.0625 = fieldNorm(doc=4373)
          0.12923586 = weight(abstract_txt:matching in 4373) [ClassicSimilarity], result of:
            0.12923586 = score(doc=4373,freq=2.0), product of:
              0.24135342 = queryWeight, product of:
                3.2373652 = boost
                6.058074 = idf(docFreq=274, maxDocs=43254)
                0.012306292 = queryNorm
              0.53546315 = fieldWeight in 4373, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                6.058074 = idf(docFreq=274, maxDocs=43254)
                0.0625 = fieldNorm(doc=4373)
          0.21371914 = weight(abstract_txt:string in 4373) [ClassicSimilarity], result of:
            0.21371914 = score(doc=4373,freq=2.0), product of:
              0.3375155 = queryWeight, product of:
                3.8283517 = boost
                7.1639853 = idf(docFreq=90, maxDocs=43254)
                0.012306292 = queryNorm
              0.6332128 = fieldWeight in 4373, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                7.1639853 = idf(docFreq=90, maxDocs=43254)
                0.0625 = fieldNorm(doc=4373)
          0.20021087 = weight(abstract_txt:approximate in 4373) [ClassicSimilarity], result of:
            0.20021087 = score(doc=4373,freq=1.0), product of:
              0.4071301 = queryWeight, product of:
                4.2046666 = boost
                7.8681827 = idf(docFreq=44, maxDocs=43254)
                0.012306292 = queryNorm
              0.49176142 = fieldWeight in 4373, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.8681827 = idf(docFreq=44, maxDocs=43254)
                0.0625 = fieldNorm(doc=4373)
          0.12368943 = weight(abstract_txt:query in 4373) [ClassicSimilarity], result of:
            0.12368943 = score(doc=4373,freq=2.0), product of:
              0.29532248 = queryWeight, product of:
                5.0644026 = boost
                4.738502 = idf(docFreq=1028, maxDocs=43254)
                0.012306292 = queryNorm
              0.41882837 = fieldWeight in 4373, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.738502 = idf(docFreq=1028, maxDocs=43254)
                0.0625 = fieldNorm(doc=4373)
        0.32 = coord(8/25)
    
  5. Airio, E.; Kettunen, K.: Does dictionary based bilingual retrieval work in a non-normalized index? (2009) 0.26
    0.25737557 = sum of:
      0.25737557 = product of:
        0.71493214 = sum of:
          0.09766225 = weight(abstract_txt:finnish in 689) [ClassicSimilarity], result of:
            0.09766225 = score(doc=689,freq=4.0), product of:
              0.10011963 = queryWeight, product of:
                1.0425445 = boost
                7.803644 = idf(docFreq=47, maxDocs=43254)
                0.012306292 = queryNorm
              0.9754555 = fieldWeight in 689, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                7.803644 = idf(docFreq=47, maxDocs=43254)
                0.0625 = fieldNorm(doc=689)
          0.022306789 = weight(abstract_txt:retrieval in 689) [ClassicSimilarity], result of:
            0.022306789 = score(doc=689,freq=3.0), product of:
              0.05938537 = queryWeight, product of:
                1.3907061 = boost
                3.4699 = idf(docFreq=3658, maxDocs=43254)
                0.012306292 = queryNorm
              0.37562767 = fieldWeight in 689, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                3.4699 = idf(docFreq=3658, maxDocs=43254)
                0.0625 = fieldNorm(doc=689)
          0.015399282 = weight(abstract_txt:were in 689) [ClassicSimilarity], result of:
            0.015399282 = score(doc=689,freq=1.0), product of:
              0.06690042 = queryWeight, product of:
                1.4760804 = boost
                3.6829145 = idf(docFreq=2956, maxDocs=43254)
                0.012306292 = queryNorm
              0.23018216 = fieldWeight in 689, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.6829145 = idf(docFreq=2956, maxDocs=43254)
                0.0625 = fieldNorm(doc=689)
          0.032879066 = weight(abstract_txt:index in 689) [ClassicSimilarity], result of:
            0.032879066 = score(doc=689,freq=1.0), product of:
              0.11092808 = queryWeight, product of:
                1.9007121 = boost
                4.7423973 = idf(docFreq=1024, maxDocs=43254)
                0.012306292 = queryNorm
              0.29639983 = fieldWeight in 689, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.7423973 = idf(docFreq=1024, maxDocs=43254)
                0.0625 = fieldNorm(doc=689)
          0.050799258 = weight(abstract_txt:forms in 689) [ClassicSimilarity], result of:
            0.050799258 = score(doc=689,freq=1.0), product of:
              0.14825185 = queryWeight, product of:
                2.197332 = boost
                5.4824824 = idf(docFreq=488, maxDocs=43254)
                0.012306292 = queryNorm
              0.34265515 = fieldWeight in 689, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.4824824 = idf(docFreq=488, maxDocs=43254)
                0.0625 = fieldNorm(doc=689)
          0.053168833 = weight(abstract_txt:test in 689) [ClassicSimilarity], result of:
            0.053168833 = score(doc=689,freq=1.0), product of:
              0.16820782 = queryWeight, product of:
                2.702639 = boost
                5.057442 = idf(docFreq=747, maxDocs=43254)
                0.012306292 = queryNorm
              0.31609014 = fieldWeight in 689, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.057442 = idf(docFreq=747, maxDocs=43254)
                0.0625 = fieldNorm(doc=689)
          0.091383554 = weight(abstract_txt:matching in 689) [ClassicSimilarity], result of:
            0.091383554 = score(doc=689,freq=1.0), product of:
              0.24135342 = queryWeight, product of:
                3.2373652 = boost
                6.058074 = idf(docFreq=274, maxDocs=43254)
                0.012306292 = queryNorm
              0.37862962 = fieldWeight in 689, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.058074 = idf(docFreq=274, maxDocs=43254)
                0.0625 = fieldNorm(doc=689)
          0.15112226 = weight(abstract_txt:string in 689) [ClassicSimilarity], result of:
            0.15112226 = score(doc=689,freq=1.0), product of:
              0.3375155 = queryWeight, product of:
                3.8283517 = boost
                7.1639853 = idf(docFreq=90, maxDocs=43254)
                0.012306292 = queryNorm
              0.44774908 = fieldWeight in 689, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.1639853 = idf(docFreq=90, maxDocs=43254)
                0.0625 = fieldNorm(doc=689)
          0.20021087 = weight(abstract_txt:approximate in 689) [ClassicSimilarity], result of:
            0.20021087 = score(doc=689,freq=1.0), product of:
              0.4071301 = queryWeight, product of:
                4.2046666 = boost
                7.8681827 = idf(docFreq=44, maxDocs=43254)
                0.012306292 = queryNorm
              0.49176142 = fieldWeight in 689, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.8681827 = idf(docFreq=44, maxDocs=43254)
                0.0625 = fieldNorm(doc=689)
        0.36 = coord(9/25)