Document (#40225)

Author
Järvelin, A.
Keskustalo, H.
Sormunen, E.
Saastamoinen, M.
Kettunen, K.
Title
Information retrieval from historical newspaper collections in highly inflectional languages : a query expansion approach
Source
Journal of the Association for Information Science and Technology. 67(2016) no.12, S.2928-2946
Year
2016
Abstract
The aim of the study was to test whether query expansion by approximate string matching methods is beneficial in retrieval from historical newspaper collections in a language rich with compounds and inflectional forms (Finnish). First, approximate string matching methods were used to generate lists of index words most similar to contemporary query terms in a digitized newspaper collection from the 1800s. Top index word variants were categorized to estimate the appropriate query expansion ranges in the retrieval test. Second, the effectiveness of approximate string matching methods, automatically generated inflectional forms, and their combinations were measured in a Cranfield-style test. Finally, a detailed topic-level analysis of test results was conducted. In the index of historical newspaper collection the occurrences of a word typically spread to many linguistic and historical variants along with optical character recognition (OCR) errors. All query expansion methods improved the baseline results. Extensive expansion of around 30 variants for each query word was required to achieve the highest performance improvement. Query expansion based on approximate string matching was superior to using the inflectional forms of the query words, showing that coverage of the different types of variation is more important than precision in handling one type of variation.
Content
Vgl.: http://onlinelibrary.wiley.com/doi/10.1002/asi.23379/full.
Theme
Computerlinguistik
Semantisches Umfeld in Indexierung u. Retrieval
Form
Zeitungen

Similar documents (author)

  1. Järvelin, K.; Kristensen, J.; Niemi, T.; Sormunen, E.; Keskustalo, H.: ¬A deductive data model for query expansion (1996) 4.89
    4.886536 = sum of:
      4.886536 = sum of:
        1.1198002 = weight(author_txt:järvelin in 4231) [ClassicSimilarity], result of:
          1.1198002 = score(doc=4231,freq=1.0), product of:
            0.44707635 = queryWeight, product of:
              8.015098 = idf(docFreq=37, maxDocs=42306)
              0.05577928 = queryNorm
            2.504718 = fieldWeight in 4231, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              8.015098 = idf(docFreq=37, maxDocs=42306)
              0.3125 = fieldNorm(doc=4231)
        1.7777175 = weight(author_txt:sormunen in 4231) [ClassicSimilarity], result of:
          1.7777175 = score(doc=4231,freq=1.0), product of:
            0.60841024 = queryWeight, product of:
              1.1665609 = boost
              9.3501 = idf(docFreq=9, maxDocs=42306)
              0.05577928 = queryNorm
            2.921906 = fieldWeight in 4231, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              9.3501 = idf(docFreq=9, maxDocs=42306)
              0.3125 = fieldNorm(doc=4231)
        1.9890186 = weight(author_txt:keskustalo in 4231) [ClassicSimilarity], result of:
          1.9890186 = score(doc=4231,freq=1.0), product of:
            0.6557132 = queryWeight, product of:
              1.2110612 = boost
              9.706774 = idf(docFreq=6, maxDocs=42306)
              0.05577928 = queryNorm
            3.0333667 = fieldWeight in 4231, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              9.706774 = idf(docFreq=6, maxDocs=42306)
              0.3125 = fieldNorm(doc=4231)
    
  2. Lehtokangas, R.; Keskustalo, H.; Järvelin, K.: Experiments with transitive dictionary translation and pseudo-relevance feedback using graded relevance assessments (2008) 2.49
    2.4870553 = sum of:
      2.4870553 = product of:
        3.7305827 = sum of:
          1.3437601 = weight(author_txt:järvelin in 3350) [ClassicSimilarity], result of:
            1.3437601 = score(doc=3350,freq=1.0), product of:
              0.44707635 = queryWeight, product of:
                8.015098 = idf(docFreq=37, maxDocs=42306)
                0.05577928 = queryNorm
              3.0056615 = fieldWeight in 3350, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.015098 = idf(docFreq=37, maxDocs=42306)
                0.375 = fieldNorm(doc=3350)
          2.3868225 = weight(author_txt:keskustalo in 3350) [ClassicSimilarity], result of:
            2.3868225 = score(doc=3350,freq=1.0), product of:
              0.6557132 = queryWeight, product of:
                1.2110612 = boost
                9.706774 = idf(docFreq=6, maxDocs=42306)
                0.05577928 = queryNorm
              3.6400402 = fieldWeight in 3350, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.706774 = idf(docFreq=6, maxDocs=42306)
                0.375 = fieldNorm(doc=3350)
        0.6666667 = coord(2/3)
    
  3. Pirkola, A.; Hedlund, T.; Keskustalo, H.; Järvelin, K.: Dictionary-based cross-language information retrieval : problems, methods, and research findings (2001) 2.07
    2.072546 = sum of:
      2.072546 = product of:
        3.1088188 = sum of:
          1.1198002 = weight(author_txt:järvelin in 4909) [ClassicSimilarity], result of:
            1.1198002 = score(doc=4909,freq=1.0), product of:
              0.44707635 = queryWeight, product of:
                8.015098 = idf(docFreq=37, maxDocs=42306)
                0.05577928 = queryNorm
              2.504718 = fieldWeight in 4909, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.015098 = idf(docFreq=37, maxDocs=42306)
                0.3125 = fieldNorm(doc=4909)
          1.9890186 = weight(author_txt:keskustalo in 4909) [ClassicSimilarity], result of:
            1.9890186 = score(doc=4909,freq=1.0), product of:
              0.6557132 = queryWeight, product of:
                1.2110612 = boost
                9.706774 = idf(docFreq=6, maxDocs=42306)
                0.05577928 = queryNorm
              3.0333667 = fieldWeight in 4909, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.706774 = idf(docFreq=6, maxDocs=42306)
                0.3125 = fieldNorm(doc=4909)
        0.6666667 = coord(2/3)
    
  4. Toivonen, J.; Pirkola, A.; Keskustalo, H.; Visala, K.; Järvelin, K.: Translating cross-lingual spelling variants using transformation rules (2005) 2.07
    2.072546 = sum of:
      2.072546 = product of:
        3.1088188 = sum of:
          1.1198002 = weight(author_txt:järvelin in 3053) [ClassicSimilarity], result of:
            1.1198002 = score(doc=3053,freq=1.0), product of:
              0.44707635 = queryWeight, product of:
                8.015098 = idf(docFreq=37, maxDocs=42306)
                0.05577928 = queryNorm
              2.504718 = fieldWeight in 3053, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.015098 = idf(docFreq=37, maxDocs=42306)
                0.3125 = fieldNorm(doc=3053)
          1.9890186 = weight(author_txt:keskustalo in 3053) [ClassicSimilarity], result of:
            1.9890186 = score(doc=3053,freq=1.0), product of:
              0.6557132 = queryWeight, product of:
                1.2110612 = boost
                9.706774 = idf(docFreq=6, maxDocs=42306)
                0.05577928 = queryNorm
              3.0333667 = fieldWeight in 3053, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.706774 = idf(docFreq=6, maxDocs=42306)
                0.3125 = fieldNorm(doc=3053)
        0.6666667 = coord(2/3)
    
  5. Ferro, N.; Silvello, G.; Keskustalo, H.; Pirkola, A.; Järvelin, K.: ¬The twist measure for IR evaluation : taking user's effort into account (2016) 2.07
    2.072546 = sum of:
      2.072546 = product of:
        3.1088188 = sum of:
          1.1198002 = weight(author_txt:järvelin in 4772) [ClassicSimilarity], result of:
            1.1198002 = score(doc=4772,freq=1.0), product of:
              0.44707635 = queryWeight, product of:
                8.015098 = idf(docFreq=37, maxDocs=42306)
                0.05577928 = queryNorm
              2.504718 = fieldWeight in 4772, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.015098 = idf(docFreq=37, maxDocs=42306)
                0.3125 = fieldNorm(doc=4772)
          1.9890186 = weight(author_txt:keskustalo in 4772) [ClassicSimilarity], result of:
            1.9890186 = score(doc=4772,freq=1.0), product of:
              0.6557132 = queryWeight, product of:
                1.2110612 = boost
                9.706774 = idf(docFreq=6, maxDocs=42306)
                0.05577928 = queryNorm
              3.0333667 = fieldWeight in 4772, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.706774 = idf(docFreq=6, maxDocs=42306)
                0.3125 = fieldNorm(doc=4772)
        0.6666667 = coord(2/3)
    

Similar documents (content)

  1. French, J.C.; Powell, A.L.; Schulman, E.: Using clustering strategies for creating authority files (2000) 0.33
    0.33333266 = sum of:
      0.33333266 = product of:
        1.0416646 = sum of:
          0.008432938 = weight(abstract_txt:from in 5812) [ClassicSimilarity], result of:
            0.008432938 = score(doc=5812,freq=1.0), product of:
              0.038593605 = queryWeight, product of:
                1.1000745 = boost
                2.796878 = idf(docFreq=7014, maxDocs=42306)
                0.012543527 = queryNorm
              0.2185061 = fieldWeight in 5812, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                2.796878 = idf(docFreq=7014, maxDocs=42306)
                0.078125 = fieldNorm(doc=5812)
          0.016001767 = weight(abstract_txt:retrieval in 5812) [ClassicSimilarity], result of:
            0.016001767 = score(doc=5812,freq=1.0), product of:
              0.059152715 = queryWeight, product of:
                1.3619206 = boost
                3.4626071 = idf(docFreq=3604, maxDocs=42306)
                0.012543527 = queryNorm
              0.2705162 = fieldWeight in 5812, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.4626071 = idf(docFreq=3604, maxDocs=42306)
                0.078125 = fieldNorm(doc=5812)
          0.06274991 = weight(abstract_txt:word in 5812) [ClassicSimilarity], result of:
            0.06274991 = score(doc=5812,freq=1.0), product of:
              0.14709733 = queryWeight, product of:
                2.1476665 = boost
                5.460322 = idf(docFreq=488, maxDocs=42306)
                0.012543527 = queryNorm
              0.42658764 = fieldWeight in 5812, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.460322 = idf(docFreq=488, maxDocs=42306)
                0.078125 = fieldNorm(doc=5812)
          0.090688 = weight(abstract_txt:forms in 5812) [ClassicSimilarity], result of:
            0.090688 = score(doc=5812,freq=2.0), product of:
              0.14924024 = queryWeight, product of:
                2.1632535 = boost
                5.4999514 = idf(docFreq=469, maxDocs=42306)
                0.012543527 = queryNorm
              0.6076645 = fieldWeight in 5812, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.4999514 = idf(docFreq=469, maxDocs=42306)
                0.078125 = fieldNorm(doc=5812)
          0.1622052 = weight(abstract_txt:variants in 5812) [ClassicSimilarity], result of:
            0.1622052 = score(doc=5812,freq=1.0), product of:
              0.2770592 = queryWeight, product of:
                2.947479 = boost
                7.493801 = idf(docFreq=63, maxDocs=42306)
                0.012543527 = queryNorm
              0.5854532 = fieldWeight in 5812, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.493801 = idf(docFreq=63, maxDocs=42306)
                0.078125 = fieldNorm(doc=5812)
          0.16158238 = weight(abstract_txt:matching in 5812) [ClassicSimilarity], result of:
            0.16158238 = score(doc=5812,freq=2.0), product of:
              0.24141355 = queryWeight, product of:
                3.176981 = boost
                6.057973 = idf(docFreq=268, maxDocs=42306)
                0.012543527 = queryNorm
              0.6693178 = fieldWeight in 5812, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                6.057973 = idf(docFreq=268, maxDocs=42306)
                0.078125 = fieldNorm(doc=5812)
          0.18896112 = weight(abstract_txt:string in 5812) [ClassicSimilarity], result of:
            0.18896112 = score(doc=5812,freq=1.0), product of:
              0.3376167 = queryWeight, product of:
                3.7570395 = boost
                7.1640477 = idf(docFreq=88, maxDocs=42306)
                0.012543527 = queryNorm
              0.55969125 = fieldWeight in 5812, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.1640477 = idf(docFreq=88, maxDocs=42306)
                0.078125 = fieldNorm(doc=5812)
          0.35104316 = weight(abstract_txt:approximate in 5812) [ClassicSimilarity], result of:
            0.35104316 = score(doc=5812,freq=2.0), product of:
              0.40495428 = queryWeight, product of:
                4.114687 = boost
                7.8460217 = idf(docFreq=44, maxDocs=42306)
                0.012543527 = queryNorm
              0.8668711 = fieldWeight in 5812, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                7.8460217 = idf(docFreq=44, maxDocs=42306)
                0.078125 = fieldNorm(doc=5812)
        0.32 = coord(8/25)
    
  2. Galvez, C.; Moya-Anegón, F.: Approximate personal name-matching through finite-state graphs (2007) 0.33
    0.32996476 = sum of:
      0.32996476 = product of:
        0.9165687 = sum of:
          0.011685022 = weight(abstract_txt:from in 2615) [ClassicSimilarity], result of:
            0.011685022 = score(doc=2615,freq=3.0), product of:
              0.038593605 = queryWeight, product of:
                1.1000745 = boost
                2.796878 = idf(docFreq=7014, maxDocs=42306)
                0.012543527 = queryNorm
              0.30277094 = fieldWeight in 2615, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                2.796878 = idf(docFreq=7014, maxDocs=42306)
                0.0625 = fieldNorm(doc=2615)
          0.012801413 = weight(abstract_txt:retrieval in 2615) [ClassicSimilarity], result of:
            0.012801413 = score(doc=2615,freq=1.0), product of:
              0.059152715 = queryWeight, product of:
                1.3619206 = boost
                3.4626071 = idf(docFreq=3604, maxDocs=42306)
                0.012543527 = queryNorm
              0.21641295 = fieldWeight in 2615, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.4626071 = idf(docFreq=3604, maxDocs=42306)
                0.0625 = fieldNorm(doc=2615)
          0.046293214 = weight(abstract_txt:index in 2615) [ClassicSimilarity], result of:
            0.046293214 = score(doc=2615,freq=2.0), product of:
              0.11061252 = queryWeight, product of:
                1.8623726 = boost
                4.7349787 = idf(docFreq=1009, maxDocs=42306)
                0.012543527 = queryNorm
              0.41851693 = fieldWeight in 2615, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.7349787 = idf(docFreq=1009, maxDocs=42306)
                0.0625 = fieldNorm(doc=2615)
          0.10260176 = weight(abstract_txt:forms in 2615) [ClassicSimilarity], result of:
            0.10260176 = score(doc=2615,freq=4.0), product of:
              0.14924024 = queryWeight, product of:
                2.1632535 = boost
                5.4999514 = idf(docFreq=469, maxDocs=42306)
                0.012543527 = queryNorm
              0.6874939 = fieldWeight in 2615, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                5.4999514 = idf(docFreq=469, maxDocs=42306)
                0.0625 = fieldNorm(doc=2615)
          0.042505253 = weight(abstract_txt:methods in 2615) [ClassicSimilarity], result of:
            0.042505253 = score(doc=2615,freq=2.0), product of:
              0.11500959 = queryWeight, product of:
                2.192809 = boost
                4.181321 = idf(docFreq=1756, maxDocs=42306)
                0.012543527 = queryNorm
              0.36958006 = fieldWeight in 2615, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.181321 = idf(docFreq=1756, maxDocs=42306)
                0.0625 = fieldNorm(doc=2615)
          0.2595283 = weight(abstract_txt:variants in 2615) [ClassicSimilarity], result of:
            0.2595283 = score(doc=2615,freq=4.0), product of:
              0.2770592 = queryWeight, product of:
                2.947479 = boost
                7.493801 = idf(docFreq=63, maxDocs=42306)
                0.012543527 = queryNorm
              0.93672514 = fieldWeight in 2615, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                7.493801 = idf(docFreq=63, maxDocs=42306)
                0.0625 = fieldNorm(doc=2615)
          0.091404796 = weight(abstract_txt:matching in 2615) [ClassicSimilarity], result of:
            0.091404796 = score(doc=2615,freq=1.0), product of:
              0.24141355 = queryWeight, product of:
                3.176981 = boost
                6.057973 = idf(docFreq=268, maxDocs=42306)
                0.012543527 = queryNorm
              0.3786233 = fieldWeight in 2615, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.057973 = idf(docFreq=268, maxDocs=42306)
                0.0625 = fieldNorm(doc=2615)
          0.15116888 = weight(abstract_txt:string in 2615) [ClassicSimilarity], result of:
            0.15116888 = score(doc=2615,freq=1.0), product of:
              0.3376167 = queryWeight, product of:
                3.7570395 = boost
                7.1640477 = idf(docFreq=88, maxDocs=42306)
                0.012543527 = queryNorm
              0.44775298 = fieldWeight in 2615, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.1640477 = idf(docFreq=88, maxDocs=42306)
                0.0625 = fieldNorm(doc=2615)
          0.19858001 = weight(abstract_txt:approximate in 2615) [ClassicSimilarity], result of:
            0.19858001 = score(doc=2615,freq=1.0), product of:
              0.40495428 = queryWeight, product of:
                4.114687 = boost
                7.8460217 = idf(docFreq=44, maxDocs=42306)
                0.012543527 = queryNorm
              0.49037635 = fieldWeight in 2615, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.8460217 = idf(docFreq=44, maxDocs=42306)
                0.0625 = fieldNorm(doc=2615)
        0.36 = coord(9/25)
    
  3. Pirkola, A.; Puolamäki, D.; Järvelin, K.: Applying query structuring in cross-language retrieval (2003) 0.30
    0.30284932 = sum of:
      0.30284932 = product of:
        0.94640416 = sum of:
          0.10658382 = weight(abstract_txt:finnish in 3075) [ClassicSimilarity], result of:
            0.10658382 = score(doc=3075,freq=3.0), product of:
              0.10067217 = queryWeight, product of:
                1.0257902 = boost
                7.824043 = idf(docFreq=45, maxDocs=42306)
                0.012543527 = queryNorm
              1.0587218 = fieldWeight in 3075, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                7.824043 = idf(docFreq=45, maxDocs=42306)
                0.078125 = fieldNorm(doc=3075)
          0.027715873 = weight(abstract_txt:retrieval in 3075) [ClassicSimilarity], result of:
            0.027715873 = score(doc=3075,freq=3.0), product of:
              0.059152715 = queryWeight, product of:
                1.3619206 = boost
                3.4626071 = idf(docFreq=3604, maxDocs=42306)
                0.012543527 = queryNorm
              0.46854776 = fieldWeight in 3075, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                3.4626071 = idf(docFreq=3604, maxDocs=42306)
                0.078125 = fieldNorm(doc=3075)
          0.039105467 = weight(abstract_txt:were in 3075) [ClassicSimilarity], result of:
            0.039105467 = score(doc=3075,freq=4.0), product of:
              0.06760846 = queryWeight, product of:
                1.456012 = boost
                3.7018294 = idf(docFreq=2837, maxDocs=42306)
                0.012543527 = queryNorm
              0.57841086 = fieldWeight in 3075, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                3.7018294 = idf(docFreq=2837, maxDocs=42306)
                0.078125 = fieldNorm(doc=3075)
          0.09460069 = weight(abstract_txt:test in 3075) [ClassicSimilarity], result of:
            0.09460069 = score(doc=3075,freq=2.0), product of:
              0.16895142 = queryWeight, product of:
                2.6577537 = boost
                5.067893 = idf(docFreq=723, maxDocs=42306)
                0.012543527 = queryNorm
              0.55992836 = fieldWeight in 3075, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.067893 = idf(docFreq=723, maxDocs=42306)
                0.078125 = fieldNorm(doc=3075)
          0.1622052 = weight(abstract_txt:variants in 3075) [ClassicSimilarity], result of:
            0.1622052 = score(doc=3075,freq=1.0), product of:
              0.2770592 = queryWeight, product of:
                2.947479 = boost
                7.493801 = idf(docFreq=63, maxDocs=42306)
                0.012543527 = queryNorm
              0.5854532 = fieldWeight in 3075, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.493801 = idf(docFreq=63, maxDocs=42306)
                0.078125 = fieldNorm(doc=3075)
          0.114255995 = weight(abstract_txt:matching in 3075) [ClassicSimilarity], result of:
            0.114255995 = score(doc=3075,freq=1.0), product of:
              0.24141355 = queryWeight, product of:
                3.176981 = boost
                6.057973 = idf(docFreq=268, maxDocs=42306)
                0.012543527 = queryNorm
              0.47327912 = fieldWeight in 3075, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.057973 = idf(docFreq=268, maxDocs=42306)
                0.078125 = fieldNorm(doc=3075)
          0.1838456 = weight(abstract_txt:newspaper in 3075) [ClassicSimilarity], result of:
            0.1838456 = score(doc=3075,freq=1.0), product of:
              0.3314956 = queryWeight, product of:
                3.7228255 = boost
                7.0988073 = idf(docFreq=94, maxDocs=42306)
                0.012543527 = queryNorm
              0.55459434 = fieldWeight in 3075, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.0988073 = idf(docFreq=94, maxDocs=42306)
                0.078125 = fieldNorm(doc=3075)
          0.21809147 = weight(abstract_txt:query in 3075) [ClassicSimilarity], result of:
            0.21809147 = score(doc=3075,freq=4.0), product of:
              0.2948434 = queryWeight, product of:
                4.9652886 = boost
                4.733989 = idf(docFreq=1010, maxDocs=42306)
                0.012543527 = queryNorm
              0.7396858 = fieldWeight in 3075, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                4.733989 = idf(docFreq=1010, maxDocs=42306)
                0.078125 = fieldNorm(doc=3075)
        0.32 = coord(8/25)
    
  4. Bellaachia, A.; Amor-Tijani, G.: Proper nouns in English-Arabic cross language information retrieval (2008) 0.29
    0.28645295 = sum of:
      0.28645295 = product of:
        0.89516544 = sum of:
          0.012801413 = weight(abstract_txt:retrieval in 192) [ClassicSimilarity], result of:
            0.012801413 = score(doc=192,freq=1.0), product of:
              0.059152715 = queryWeight, product of:
                1.3619206 = boost
                3.4626071 = idf(docFreq=3604, maxDocs=42306)
                0.012543527 = queryNorm
              0.21641295 = fieldWeight in 192, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.4626071 = idf(docFreq=3604, maxDocs=42306)
                0.0625 = fieldNorm(doc=192)
          0.054863434 = weight(abstract_txt:words in 192) [ClassicSimilarity], result of:
            0.054863434 = score(doc=192,freq=3.0), product of:
              0.09453382 = queryWeight, product of:
                1.405764 = boost
                5.361115 = idf(docFreq=539, maxDocs=42306)
                0.012543527 = queryNorm
              0.58035773 = fieldWeight in 192, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                5.361115 = idf(docFreq=539, maxDocs=42306)
                0.0625 = fieldNorm(doc=192)
          0.032734245 = weight(abstract_txt:index in 192) [ClassicSimilarity], result of:
            0.032734245 = score(doc=192,freq=1.0), product of:
              0.11061252 = queryWeight, product of:
                1.8623726 = boost
                4.7349787 = idf(docFreq=1009, maxDocs=42306)
                0.012543527 = queryNorm
              0.29593617 = fieldWeight in 192, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.7349787 = idf(docFreq=1009, maxDocs=42306)
                0.0625 = fieldNorm(doc=192)
          0.12976415 = weight(abstract_txt:variants in 192) [ClassicSimilarity], result of:
            0.12976415 = score(doc=192,freq=1.0), product of:
              0.2770592 = queryWeight, product of:
                2.947479 = boost
                7.493801 = idf(docFreq=63, maxDocs=42306)
                0.012543527 = queryNorm
              0.46836257 = fieldWeight in 192, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.493801 = idf(docFreq=63, maxDocs=42306)
                0.0625 = fieldNorm(doc=192)
          0.1292659 = weight(abstract_txt:matching in 192) [ClassicSimilarity], result of:
            0.1292659 = score(doc=192,freq=2.0), product of:
              0.24141355 = queryWeight, product of:
                3.176981 = boost
                6.057973 = idf(docFreq=268, maxDocs=42306)
                0.012543527 = queryNorm
              0.5354542 = fieldWeight in 192, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                6.057973 = idf(docFreq=268, maxDocs=42306)
                0.0625 = fieldNorm(doc=192)
          0.2137851 = weight(abstract_txt:string in 192) [ClassicSimilarity], result of:
            0.2137851 = score(doc=192,freq=2.0), product of:
              0.3376167 = queryWeight, product of:
                3.7570395 = boost
                7.1640477 = idf(docFreq=88, maxDocs=42306)
                0.012543527 = queryNorm
              0.63321835 = fieldWeight in 192, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                7.1640477 = idf(docFreq=88, maxDocs=42306)
                0.0625 = fieldNorm(doc=192)
          0.19858001 = weight(abstract_txt:approximate in 192) [ClassicSimilarity], result of:
            0.19858001 = score(doc=192,freq=1.0), product of:
              0.40495428 = queryWeight, product of:
                4.114687 = boost
                7.8460217 = idf(docFreq=44, maxDocs=42306)
                0.012543527 = queryNorm
              0.49037635 = fieldWeight in 192, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.8460217 = idf(docFreq=44, maxDocs=42306)
                0.0625 = fieldNorm(doc=192)
          0.12337116 = weight(abstract_txt:query in 192) [ClassicSimilarity], result of:
            0.12337116 = score(doc=192,freq=2.0), product of:
              0.2948434 = queryWeight, product of:
                4.9652886 = boost
                4.733989 = idf(docFreq=1010, maxDocs=42306)
                0.012543527 = queryNorm
              0.41842943 = fieldWeight in 192, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.733989 = idf(docFreq=1010, maxDocs=42306)
                0.0625 = fieldNorm(doc=192)
        0.32 = coord(8/25)
    
  5. Airio, E.; Kettunen, K.: Does dictionary based bilingual retrieval work in a non-normalized index? (2009) 0.26
    0.2573913 = sum of:
      0.2573913 = product of:
        0.71497583 = sum of:
          0.098457925 = weight(abstract_txt:finnish in 1225) [ClassicSimilarity], result of:
            0.098457925 = score(doc=1225,freq=4.0), product of:
              0.10067217 = queryWeight, product of:
                1.0257902 = boost
                7.824043 = idf(docFreq=45, maxDocs=42306)
                0.012543527 = queryNorm
              0.97800535 = fieldWeight in 1225, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                7.824043 = idf(docFreq=45, maxDocs=42306)
                0.0625 = fieldNorm(doc=1225)
          0.022172697 = weight(abstract_txt:retrieval in 1225) [ClassicSimilarity], result of:
            0.022172697 = score(doc=1225,freq=3.0), product of:
              0.059152715 = queryWeight, product of:
                1.3619206 = boost
                3.4626071 = idf(docFreq=3604, maxDocs=42306)
                0.012543527 = queryNorm
              0.3748382 = fieldWeight in 1225, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                3.4626071 = idf(docFreq=3604, maxDocs=42306)
                0.0625 = fieldNorm(doc=1225)
          0.015642187 = weight(abstract_txt:were in 1225) [ClassicSimilarity], result of:
            0.015642187 = score(doc=1225,freq=1.0), product of:
              0.06760846 = queryWeight, product of:
                1.456012 = boost
                3.7018294 = idf(docFreq=2837, maxDocs=42306)
                0.012543527 = queryNorm
              0.23136434 = fieldWeight in 1225, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.7018294 = idf(docFreq=2837, maxDocs=42306)
                0.0625 = fieldNorm(doc=1225)
          0.032734245 = weight(abstract_txt:index in 1225) [ClassicSimilarity], result of:
            0.032734245 = score(doc=1225,freq=1.0), product of:
              0.11061252 = queryWeight, product of:
                1.8623726 = boost
                4.7349787 = idf(docFreq=1009, maxDocs=42306)
                0.012543527 = queryNorm
              0.29593617 = fieldWeight in 1225, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.7349787 = idf(docFreq=1009, maxDocs=42306)
                0.0625 = fieldNorm(doc=1225)
          0.05130088 = weight(abstract_txt:forms in 1225) [ClassicSimilarity], result of:
            0.05130088 = score(doc=1225,freq=1.0), product of:
              0.14924024 = queryWeight, product of:
                2.1632535 = boost
                5.4999514 = idf(docFreq=469, maxDocs=42306)
                0.012543527 = queryNorm
              0.34374696 = fieldWeight in 1225, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.4999514 = idf(docFreq=469, maxDocs=42306)
                0.0625 = fieldNorm(doc=1225)
          0.053514235 = weight(abstract_txt:test in 1225) [ClassicSimilarity], result of:
            0.053514235 = score(doc=1225,freq=1.0), product of:
              0.16895142 = queryWeight, product of:
                2.6577537 = boost
                5.067893 = idf(docFreq=723, maxDocs=42306)
                0.012543527 = queryNorm
              0.3167433 = fieldWeight in 1225, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.067893 = idf(docFreq=723, maxDocs=42306)
                0.0625 = fieldNorm(doc=1225)
          0.091404796 = weight(abstract_txt:matching in 1225) [ClassicSimilarity], result of:
            0.091404796 = score(doc=1225,freq=1.0), product of:
              0.24141355 = queryWeight, product of:
                3.176981 = boost
                6.057973 = idf(docFreq=268, maxDocs=42306)
                0.012543527 = queryNorm
              0.3786233 = fieldWeight in 1225, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.057973 = idf(docFreq=268, maxDocs=42306)
                0.0625 = fieldNorm(doc=1225)
          0.15116888 = weight(abstract_txt:string in 1225) [ClassicSimilarity], result of:
            0.15116888 = score(doc=1225,freq=1.0), product of:
              0.3376167 = queryWeight, product of:
                3.7570395 = boost
                7.1640477 = idf(docFreq=88, maxDocs=42306)
                0.012543527 = queryNorm
              0.44775298 = fieldWeight in 1225, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.1640477 = idf(docFreq=88, maxDocs=42306)
                0.0625 = fieldNorm(doc=1225)
          0.19858001 = weight(abstract_txt:approximate in 1225) [ClassicSimilarity], result of:
            0.19858001 = score(doc=1225,freq=1.0), product of:
              0.40495428 = queryWeight, product of:
                4.114687 = boost
                7.8460217 = idf(docFreq=44, maxDocs=42306)
                0.012543527 = queryNorm
              0.49037635 = fieldWeight in 1225, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.8460217 = idf(docFreq=44, maxDocs=42306)
                0.0625 = fieldNorm(doc=1225)
        0.36 = coord(9/25)