Document (#38285)

Author
Dang, E.K.F.
Luk, R.W.P.
Allan, J.
Title
Beyond bag-of-words : bigram-enhanced context-dependent term weights
Source
Journal of the Association for Information Science and Technology. 65(2014) no.6, S.1134-1148
Year
2014
Abstract
While term independence is a widely held assumption in most of the established information retrieval approaches, it is clearly not true and various works in the past have investigated a relaxation of the assumption. One approach is to use n-grams in document representation instead of unigrams. However, the majority of early works on n-grams obtained only modest performance improvement. On the other hand, the use of information based on supporting terms or "contexts" of queries has been found to be promising. In particular, recent studies showed that using new context-dependent term weights improved the performance of relevance feedback (RF) retrieval compared with using traditional bag-of-words BM25 term weights. Calculation of the new term weights requires an estimation of the local probability of relevance of each query term occurrence. In previous studies, the estimation of this probability was based on unigrams that occur in the neighborhood of a query term. We explore an integration of the n-gram and context approaches by computing context-dependent term weights based on a mixture of unigrams and bigrams. Extensive experiments are performed using the title queries of the Text Retrieval Conference (TREC)-6, TREC-7, TREC-8, and TREC-2005 collections, for RF with relevance judgment of either the top 10 or top 20 documents of an initial retrieval. We identify some crucial elements needed in the use of bigrams in our methods, such as proper inverse document frequency (IDF) weighting of the bigrams and noise reduction by pruning bigrams with large document frequency values. We show that enhancing context-dependent term weights with bigrams is effective in further improving retrieval performance.
Theme
Retrievalalgorithmen
Object
Bigrams

Similar documents (author)

  1. Dang, E.K.F.; Luk, R.W.P.; Allan, J.: ¬A context-dependent relevance model (2016) 6.04
    6.0361447 = sum of:
      6.0361447 = sum of:
        1.7193992 = weight(author_txt:allan in 4779) [ClassicSimilarity], result of:
          1.7193992 = score(doc=4779,freq=1.0), product of:
            0.5192788 = queryWeight, product of:
              8.829678 = idf(docFreq=16, maxDocs=42740)
              0.05881062 = queryNorm
            3.311129 = fieldWeight in 4779, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              8.829678 = idf(docFreq=16, maxDocs=42740)
              0.375 = fieldNorm(doc=4779)
        2.1183403 = weight(author_txt:r.w.p in 4779) [ClassicSimilarity], result of:
          2.1183403 = score(doc=4779,freq=1.0), product of:
            0.59677863 = queryWeight, product of:
              1.0720285 = boost
              9.465666 = idf(docFreq=8, maxDocs=42740)
              0.05881062 = queryNorm
            3.5496247 = fieldWeight in 4779, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              9.465666 = idf(docFreq=8, maxDocs=42740)
              0.375 = fieldNorm(doc=4779)
        2.1984053 = weight(author_txt:dang in 4779) [ClassicSimilarity], result of:
          2.1984053 = score(doc=4779,freq=1.0), product of:
            0.61172277 = queryWeight, product of:
              1.085368 = boost
              9.583449 = idf(docFreq=7, maxDocs=42740)
              0.05881062 = queryNorm
            3.5937934 = fieldWeight in 4779, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              9.583449 = idf(docFreq=7, maxDocs=42740)
              0.375 = fieldNorm(doc=4779)
    
  2. Dang, E.K.F.; Luk, R.W.P.; Allan, J.; Ho, K.S.; Chung, K.F.L.; Lee, D.L.: ¬A new context-dependent term weight computed by boost and discount using relevance information (2010) 4.02
    4.0240965 = sum of:
      4.0240965 = sum of:
        1.1462661 = weight(author_txt:allan in 1121) [ClassicSimilarity], result of:
          1.1462661 = score(doc=1121,freq=1.0), product of:
            0.5192788 = queryWeight, product of:
              8.829678 = idf(docFreq=16, maxDocs=42740)
              0.05881062 = queryNorm
            2.2074194 = fieldWeight in 1121, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              8.829678 = idf(docFreq=16, maxDocs=42740)
              0.25 = fieldNorm(doc=1121)
        1.4122268 = weight(author_txt:r.w.p in 1121) [ClassicSimilarity], result of:
          1.4122268 = score(doc=1121,freq=1.0), product of:
            0.59677863 = queryWeight, product of:
              1.0720285 = boost
              9.465666 = idf(docFreq=8, maxDocs=42740)
              0.05881062 = queryNorm
            2.3664165 = fieldWeight in 1121, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              9.465666 = idf(docFreq=8, maxDocs=42740)
              0.25 = fieldNorm(doc=1121)
        1.4656036 = weight(author_txt:dang in 1121) [ClassicSimilarity], result of:
          1.4656036 = score(doc=1121,freq=1.0), product of:
            0.61172277 = queryWeight, product of:
              1.085368 = boost
              9.583449 = idf(docFreq=7, maxDocs=42740)
              0.05881062 = queryNorm
            2.3958623 = fieldWeight in 1121, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              9.583449 = idf(docFreq=7, maxDocs=42740)
              0.25 = fieldNorm(doc=1121)
    
  3. Dang, E.K.F.; Luk, R.W.P.; Ho, K.S.; Chan, S.C.F.; Lee, D.L.: ¬A new measure of clustering effectiveness : algorithms and experimental studies (2008) 2.40
    2.3981922 = sum of:
      2.3981922 = product of:
        3.5972881 = sum of:
          1.7652836 = weight(author_txt:r.w.p in 3368) [ClassicSimilarity], result of:
            1.7652836 = score(doc=3368,freq=1.0), product of:
              0.59677863 = queryWeight, product of:
                1.0720285 = boost
                9.465666 = idf(docFreq=8, maxDocs=42740)
                0.05881062 = queryNorm
              2.9580207 = fieldWeight in 3368, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.465666 = idf(docFreq=8, maxDocs=42740)
                0.3125 = fieldNorm(doc=3368)
          1.8320044 = weight(author_txt:dang in 3368) [ClassicSimilarity], result of:
            1.8320044 = score(doc=3368,freq=1.0), product of:
              0.61172277 = queryWeight, product of:
                1.085368 = boost
                9.583449 = idf(docFreq=7, maxDocs=42740)
                0.05881062 = queryNorm
              2.994828 = fieldWeight in 3368, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.583449 = idf(docFreq=7, maxDocs=42740)
                0.3125 = fieldNorm(doc=3368)
        0.6666667 = coord(2/3)
    
  4. Allan, A.: Library malpractice suit : could it happen to you? (1976) 0.96
    0.95522183 = sum of:
      0.95522183 = product of:
        2.8656654 = sum of:
          2.8656654 = weight(author_txt:allan in 3653) [ClassicSimilarity], result of:
            2.8656654 = score(doc=3653,freq=1.0), product of:
              0.5192788 = queryWeight, product of:
                8.829678 = idf(docFreq=16, maxDocs=42740)
                0.05881062 = queryNorm
              5.5185485 = fieldWeight in 3653, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.829678 = idf(docFreq=16, maxDocs=42740)
                0.625 = fieldNorm(doc=3653)
        0.33333334 = coord(1/3)
    
  5. Allan, J.: Building hypertext using information retrieval (1997) 0.96
    0.95522183 = sum of:
      0.95522183 = product of:
        2.8656654 = sum of:
          2.8656654 = weight(author_txt:allan in 1149) [ClassicSimilarity], result of:
            2.8656654 = score(doc=1149,freq=1.0), product of:
              0.5192788 = queryWeight, product of:
                8.829678 = idf(docFreq=16, maxDocs=42740)
                0.05881062 = queryNorm
              5.5185485 = fieldWeight in 1149, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.829678 = idf(docFreq=16, maxDocs=42740)
                0.625 = fieldNorm(doc=1149)
        0.33333334 = coord(1/3)
    

Similar documents (content)

  1. Dang, E.K.F.; Luk, R.W.P.; Allan, J.; Ho, K.S.; Chung, K.F.L.; Lee, D.L.: ¬A new context-dependent term weight computed by boost and discount using relevance information (2010) 0.68
    0.6842551 = sum of:
      0.6842551 = product of:
        1.3158753 = sum of:
          0.018137846 = weight(abstract_txt:query in 1121) [ClassicSimilarity], result of:
            0.018137846 = score(doc=1121,freq=1.0), product of:
              0.061310507 = queryWeight, product of:
                1.0113716 = boost
                4.7333736 = idf(docFreq=1021, maxDocs=42740)
                0.012807176 = queryNorm
              0.29583585 = fieldWeight in 1121, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.7333736 = idf(docFreq=1021, maxDocs=42740)
                0.0625 = fieldNorm(doc=1121)
          0.010917044 = weight(abstract_txt:with in 1121) [ClassicSimilarity], result of:
            0.010917044 = score(doc=1121,freq=4.0), product of:
              0.034690015 = queryWeight, product of:
                1.0758718 = boost
                2.5176222 = idf(docFreq=9369, maxDocs=42740)
                0.012807176 = queryNorm
              0.31470278 = fieldWeight in 1121, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                2.5176222 = idf(docFreq=9369, maxDocs=42740)
                0.0625 = fieldNorm(doc=1121)
          0.022543669 = weight(abstract_txt:queries in 1121) [ClassicSimilarity], result of:
            0.022543669 = score(doc=1121,freq=1.0), product of:
              0.070875175 = queryWeight, product of:
                1.0874026 = boost
                5.0892105 = idf(docFreq=715, maxDocs=42740)
                0.012807176 = queryNorm
              0.31807566 = fieldWeight in 1121, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.0892105 = idf(docFreq=715, maxDocs=42740)
                0.0625 = fieldNorm(doc=1121)
          0.018718595 = weight(abstract_txt:using in 1121) [ClassicSimilarity], result of:
            0.018718595 = score(doc=1121,freq=3.0), product of:
              0.049695447 = queryWeight, product of:
                1.1151857 = boost
                3.4794931 = idf(docFreq=3580, maxDocs=42740)
                0.012807176 = queryNorm
              0.3766662 = fieldWeight in 1121, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                3.4794931 = idf(docFreq=3580, maxDocs=42740)
                0.0625 = fieldNorm(doc=1121)
          0.051269844 = weight(abstract_txt:frequency in 1121) [ClassicSimilarity], result of:
            0.051269844 = score(doc=1121,freq=2.0), product of:
              0.09728423 = queryWeight, product of:
                1.2739855 = boost
                5.962447 = idf(docFreq=298, maxDocs=42740)
                0.012807176 = queryNorm
              0.52701086 = fieldWeight in 1121, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.962447 = idf(docFreq=298, maxDocs=42740)
                0.0625 = fieldNorm(doc=1121)
          0.034856975 = weight(abstract_txt:document in 1121) [ClassicSimilarity], result of:
            0.034856975 = score(doc=1121,freq=3.0), product of:
              0.075219005 = queryWeight, product of:
                1.3719956 = boost
                4.280766 = idf(docFreq=1606, maxDocs=42740)
                0.012807176 = queryNorm
              0.4634065 = fieldWeight in 1121, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.280766 = idf(docFreq=1606, maxDocs=42740)
                0.0625 = fieldNorm(doc=1121)
          0.053381637 = weight(abstract_txt:relevance in 1121) [ClassicSimilarity], result of:
            0.053381637 = score(doc=1121,freq=3.0), product of:
              0.099937625 = queryWeight, product of:
                1.5814426 = boost
                4.934262 = idf(docFreq=835, maxDocs=42740)
                0.012807176 = queryNorm
              0.5341495 = fieldWeight in 1121, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.934262 = idf(docFreq=835, maxDocs=42740)
                0.0625 = fieldNorm(doc=1121)
          0.017784778 = weight(abstract_txt:retrieval in 1121) [ClassicSimilarity], result of:
            0.017784778 = score(doc=1121,freq=1.0), product of:
              0.08212778 = queryWeight, product of:
                1.850795 = boost
                3.4648013 = idf(docFreq=3633, maxDocs=42740)
                0.012807176 = queryNorm
              0.21655008 = fieldWeight in 1121, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.4648013 = idf(docFreq=3633, maxDocs=42740)
                0.0625 = fieldNorm(doc=1121)
          0.071727775 = weight(abstract_txt:context in 1121) [ClassicSimilarity], result of:
            0.071727775 = score(doc=1121,freq=4.0), product of:
              0.13108793 = queryWeight, product of:
                2.33827 = boost
                4.377384 = idf(docFreq=1458, maxDocs=42740)
                0.012807176 = queryNorm
              0.547173 = fieldWeight in 1121, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                4.377384 = idf(docFreq=1458, maxDocs=42740)
                0.0625 = fieldNorm(doc=1121)
          0.12775686 = weight(abstract_txt:dependent in 1121) [ClassicSimilarity], result of:
            0.12775686 = score(doc=1121,freq=2.0), product of:
              0.22528586 = queryWeight, product of:
                2.7417336 = boost
                6.4158664 = idf(docFreq=189, maxDocs=42740)
                0.012807176 = queryNorm
              0.5670878 = fieldWeight in 1121, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                6.4158664 = idf(docFreq=189, maxDocs=42740)
                0.0625 = fieldNorm(doc=1121)
          0.102242865 = weight(abstract_txt:trec in 1121) [ClassicSimilarity], result of:
            0.102242865 = score(doc=1121,freq=1.0), product of:
              0.24466759 = queryWeight, product of:
                2.8572385 = boost
                6.6861567 = idf(docFreq=144, maxDocs=42740)
                0.012807176 = queryNorm
              0.4178848 = fieldWeight in 1121, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.6861567 = idf(docFreq=144, maxDocs=42740)
                0.0625 = fieldNorm(doc=1121)
          0.22927019 = weight(abstract_txt:term in 1121) [ClassicSimilarity], result of:
            0.22927019 = score(doc=1121,freq=7.0), product of:
              0.2871316 = queryWeight, product of:
                4.6429076 = boost
                4.8287816 = idf(docFreq=928, maxDocs=42740)
                0.012807176 = queryNorm
              0.7984847 = fieldWeight in 1121, product of:
                2.6457512 = tf(freq=7.0), with freq of:
                  7.0 = termFreq=7.0
                4.8287816 = idf(docFreq=928, maxDocs=42740)
                0.0625 = fieldNorm(doc=1121)
          0.55726725 = weight(abstract_txt:weights in 1121) [ClassicSimilarity], result of:
            0.55726725 = score(doc=1121,freq=8.0), product of:
              0.43370777 = queryWeight, product of:
                4.6591053 = boost
                7.268441 = idf(docFreq=80, maxDocs=42740)
                0.012807176 = queryNorm
              1.284891 = fieldWeight in 1121, product of:
                2.828427 = tf(freq=8.0), with freq of:
                  8.0 = termFreq=8.0
                7.268441 = idf(docFreq=80, maxDocs=42740)
                0.0625 = fieldNorm(doc=1121)
        0.52 = coord(13/25)
    
  2. Dang, E.K.F.; Luk, R.W.P.; Allan, J.: ¬A context-dependent relevance model (2016) 0.46
    0.45850354 = sum of:
      0.45850354 = product of:
        0.8817376 = sum of:
          0.03141567 = weight(abstract_txt:query in 4779) [ClassicSimilarity], result of:
            0.03141567 = score(doc=4779,freq=3.0), product of:
              0.061310507 = queryWeight, product of:
                1.0113716 = boost
                4.7333736 = idf(docFreq=1021, maxDocs=42740)
                0.012807176 = queryNorm
              0.5124027 = fieldWeight in 4779, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.7333736 = idf(docFreq=1021, maxDocs=42740)
                0.0625 = fieldNorm(doc=4779)
          0.00847643 = weight(abstract_txt:based in 4779) [ClassicSimilarity], result of:
            0.00847643 = score(doc=4779,freq=1.0), product of:
              0.042265255 = queryWeight, product of:
                1.028444 = boost
                3.2088501 = idf(docFreq=4693, maxDocs=42740)
                0.012807176 = queryNorm
              0.20055313 = fieldWeight in 4779, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.2088501 = idf(docFreq=4693, maxDocs=42740)
                0.0625 = fieldNorm(doc=4779)
          0.005458522 = weight(abstract_txt:with in 4779) [ClassicSimilarity], result of:
            0.005458522 = score(doc=4779,freq=1.0), product of:
              0.034690015 = queryWeight, product of:
                1.0758718 = boost
                2.5176222 = idf(docFreq=9369, maxDocs=42740)
                0.012807176 = queryNorm
              0.15735139 = fieldWeight in 4779, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                2.5176222 = idf(docFreq=9369, maxDocs=42740)
                0.0625 = fieldNorm(doc=4779)
          0.010807185 = weight(abstract_txt:using in 4779) [ClassicSimilarity], result of:
            0.010807185 = score(doc=4779,freq=1.0), product of:
              0.049695447 = queryWeight, product of:
                1.1151857 = boost
                3.4794931 = idf(docFreq=3580, maxDocs=42740)
                0.012807176 = queryNorm
              0.21746832 = fieldWeight in 4779, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.4794931 = idf(docFreq=3580, maxDocs=42740)
                0.0625 = fieldNorm(doc=4779)
          0.026314124 = weight(abstract_txt:words in 4779) [ClassicSimilarity], result of:
            0.026314124 = score(doc=4779,freq=1.0), product of:
              0.07857247 = queryWeight, product of:
                1.1449288 = boost
                5.358442 = idf(docFreq=546, maxDocs=42740)
                0.012807176 = queryNorm
              0.3349026 = fieldWeight in 4779, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.358442 = idf(docFreq=546, maxDocs=42740)
                0.0625 = fieldNorm(doc=4779)
          0.028460603 = weight(abstract_txt:document in 4779) [ClassicSimilarity], result of:
            0.028460603 = score(doc=4779,freq=2.0), product of:
              0.075219005 = queryWeight, product of:
                1.3719956 = boost
                4.280766 = idf(docFreq=1606, maxDocs=42740)
                0.012807176 = queryNorm
              0.37836984 = fieldWeight in 4779, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.280766 = idf(docFreq=1606, maxDocs=42740)
                0.0625 = fieldNorm(doc=4779)
          0.025672987 = weight(abstract_txt:performance in 4779) [ClassicSimilarity], result of:
            0.025672987 = score(doc=4779,freq=1.0), product of:
              0.08847606 = queryWeight, product of:
                1.4879961 = boost
                4.6426997 = idf(docFreq=1118, maxDocs=42740)
                0.012807176 = queryNorm
              0.29016873 = fieldWeight in 4779, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6426997 = idf(docFreq=1118, maxDocs=42740)
                0.0625 = fieldNorm(doc=4779)
          0.0689154 = weight(abstract_txt:relevance in 4779) [ClassicSimilarity], result of:
            0.0689154 = score(doc=4779,freq=5.0), product of:
              0.099937625 = queryWeight, product of:
                1.5814426 = boost
                4.934262 = idf(docFreq=835, maxDocs=42740)
                0.012807176 = queryNorm
              0.6895841 = fieldWeight in 4779, product of:
                2.236068 = tf(freq=5.0), with freq of:
                  5.0 = termFreq=5.0
                4.934262 = idf(docFreq=835, maxDocs=42740)
                0.0625 = fieldNorm(doc=4779)
          0.035569556 = weight(abstract_txt:retrieval in 4779) [ClassicSimilarity], result of:
            0.035569556 = score(doc=4779,freq=4.0), product of:
              0.08212778 = queryWeight, product of:
                1.850795 = boost
                3.4648013 = idf(docFreq=3633, maxDocs=42740)
                0.012807176 = queryNorm
              0.43310016 = fieldWeight in 4779, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                3.4648013 = idf(docFreq=3633, maxDocs=42740)
                0.0625 = fieldNorm(doc=4779)
          0.071727775 = weight(abstract_txt:context in 4779) [ClassicSimilarity], result of:
            0.071727775 = score(doc=4779,freq=4.0), product of:
              0.13108793 = queryWeight, product of:
                2.33827 = boost
                4.377384 = idf(docFreq=1458, maxDocs=42740)
                0.012807176 = queryNorm
              0.547173 = fieldWeight in 4779, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                4.377384 = idf(docFreq=1458, maxDocs=42740)
                0.0625 = fieldNorm(doc=4779)
          0.090337746 = weight(abstract_txt:dependent in 4779) [ClassicSimilarity], result of:
            0.090337746 = score(doc=4779,freq=1.0), product of:
              0.22528586 = queryWeight, product of:
                2.7417336 = boost
                6.4158664 = idf(docFreq=189, maxDocs=42740)
                0.012807176 = queryNorm
              0.40099165 = fieldWeight in 4779, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.4158664 = idf(docFreq=189, maxDocs=42740)
                0.0625 = fieldNorm(doc=4779)
          0.102242865 = weight(abstract_txt:trec in 4779) [ClassicSimilarity], result of:
            0.102242865 = score(doc=4779,freq=1.0), product of:
              0.24466759 = queryWeight, product of:
                2.8572385 = boost
                6.6861567 = idf(docFreq=144, maxDocs=42740)
                0.012807176 = queryNorm
              0.4178848 = fieldWeight in 4779, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.6861567 = idf(docFreq=144, maxDocs=42740)
                0.0625 = fieldNorm(doc=4779)
          0.37633875 = weight(abstract_txt:bigrams in 4779) [ClassicSimilarity], result of:
            0.37633875 = score(doc=4779,freq=1.0), product of:
              0.6283145 = queryWeight, product of:
                5.119197 = boost
                9.583449 = idf(docFreq=7, maxDocs=42740)
                0.012807176 = queryNorm
              0.5989656 = fieldWeight in 4779, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.583449 = idf(docFreq=7, maxDocs=42740)
                0.0625 = fieldNorm(doc=4779)
        0.52 = coord(13/25)
    
  3. Lhadj, L.S.; Boughanem, M.; Amrouche, K.: Enhancing information retrieval through concept-based language modeling and semantic smoothing (2016) 0.37
    0.37142104 = sum of:
      0.37142104 = product of:
        0.84413874 = sum of:
          0.014681606 = weight(abstract_txt:based in 5222) [ClassicSimilarity], result of:
            0.014681606 = score(doc=5222,freq=3.0), product of:
              0.042265255 = queryWeight, product of:
                1.028444 = boost
                3.2088501 = idf(docFreq=4693, maxDocs=42740)
                0.012807176 = queryNorm
              0.3473682 = fieldWeight in 5222, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                3.2088501 = idf(docFreq=4693, maxDocs=42740)
                0.0625 = fieldNorm(doc=5222)
          0.0094544375 = weight(abstract_txt:with in 5222) [ClassicSimilarity], result of:
            0.0094544375 = score(doc=5222,freq=3.0), product of:
              0.034690015 = queryWeight, product of:
                1.0758718 = boost
                2.5176222 = idf(docFreq=9369, maxDocs=42740)
                0.012807176 = queryNorm
              0.2725406 = fieldWeight in 5222, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                2.5176222 = idf(docFreq=9369, maxDocs=42740)
                0.0625 = fieldNorm(doc=5222)
          0.015283668 = weight(abstract_txt:using in 5222) [ClassicSimilarity], result of:
            0.015283668 = score(doc=5222,freq=2.0), product of:
              0.049695447 = queryWeight, product of:
                1.1151857 = boost
                3.4794931 = idf(docFreq=3580, maxDocs=42740)
                0.012807176 = queryNorm
              0.30754665 = fieldWeight in 5222, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.4794931 = idf(docFreq=3580, maxDocs=42740)
                0.0625 = fieldNorm(doc=5222)
          0.020124685 = weight(abstract_txt:document in 5222) [ClassicSimilarity], result of:
            0.020124685 = score(doc=5222,freq=1.0), product of:
              0.075219005 = queryWeight, product of:
                1.3719956 = boost
                4.280766 = idf(docFreq=1606, maxDocs=42740)
                0.012807176 = queryNorm
              0.26754788 = fieldWeight in 5222, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.280766 = idf(docFreq=1606, maxDocs=42740)
                0.0625 = fieldNorm(doc=5222)
          0.07538664 = weight(abstract_txt:assumption in 5222) [ClassicSimilarity], result of:
            0.07538664 = score(doc=5222,freq=2.0), product of:
              0.1257952 = queryWeight, product of:
                1.4486895 = boost
                6.7800884 = idf(docFreq=131, maxDocs=42740)
                0.012807176 = queryNorm
              0.5992808 = fieldWeight in 5222, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                6.7800884 = idf(docFreq=131, maxDocs=42740)
                0.0625 = fieldNorm(doc=5222)
          0.025672987 = weight(abstract_txt:performance in 5222) [ClassicSimilarity], result of:
            0.025672987 = score(doc=5222,freq=1.0), product of:
              0.08847606 = queryWeight, product of:
                1.4879961 = boost
                4.6426997 = idf(docFreq=1118, maxDocs=42740)
                0.012807176 = queryNorm
              0.29016873 = fieldWeight in 5222, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6426997 = idf(docFreq=1118, maxDocs=42740)
                0.0625 = fieldNorm(doc=5222)
          0.093145646 = weight(abstract_txt:grams in 5222) [ClassicSimilarity], result of:
            0.093145646 = score(doc=5222,freq=1.0), product of:
              0.18249577 = queryWeight, product of:
                1.7448965 = boost
                8.166383 = idf(docFreq=32, maxDocs=42740)
                0.012807176 = queryNorm
              0.5103989 = fieldWeight in 5222, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.166383 = idf(docFreq=32, maxDocs=42740)
                0.0625 = fieldNorm(doc=5222)
          0.025151474 = weight(abstract_txt:retrieval in 5222) [ClassicSimilarity], result of:
            0.025151474 = score(doc=5222,freq=2.0), product of:
              0.08212778 = queryWeight, product of:
                1.850795 = boost
                3.4648013 = idf(docFreq=3633, maxDocs=42740)
                0.012807176 = queryNorm
              0.30624807 = fieldWeight in 5222, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.4648013 = idf(docFreq=3633, maxDocs=42740)
                0.0625 = fieldNorm(doc=5222)
          0.102242865 = weight(abstract_txt:trec in 5222) [ClassicSimilarity], result of:
            0.102242865 = score(doc=5222,freq=1.0), product of:
              0.24466759 = queryWeight, product of:
                2.8572385 = boost
                6.6861567 = idf(docFreq=144, maxDocs=42740)
                0.012807176 = queryNorm
              0.4178848 = fieldWeight in 5222, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.6861567 = idf(docFreq=144, maxDocs=42740)
                0.0625 = fieldNorm(doc=5222)
          0.08665599 = weight(abstract_txt:term in 5222) [ClassicSimilarity], result of:
            0.08665599 = score(doc=5222,freq=1.0), product of:
              0.2871316 = queryWeight, product of:
                4.6429076 = boost
                4.8287816 = idf(docFreq=928, maxDocs=42740)
                0.012807176 = queryNorm
              0.30179885 = fieldWeight in 5222, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.8287816 = idf(docFreq=928, maxDocs=42740)
                0.0625 = fieldNorm(doc=5222)
          0.37633875 = weight(abstract_txt:bigrams in 5222) [ClassicSimilarity], result of:
            0.37633875 = score(doc=5222,freq=1.0), product of:
              0.6283145 = queryWeight, product of:
                5.119197 = boost
                9.583449 = idf(docFreq=7, maxDocs=42740)
                0.012807176 = queryNorm
              0.5989656 = fieldWeight in 5222, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.583449 = idf(docFreq=7, maxDocs=42740)
                0.0625 = fieldNorm(doc=5222)
        0.44 = coord(11/25)
    
  4. Ruthven, I.; Lalmas, M.; Rijsbergen, K. van: Combining and selecting characteristics of information use (2002) 0.36
    0.3568005 = sum of:
      0.3568005 = product of:
        0.63714373 = sum of:
          0.03599117 = weight(abstract_txt:query in 209) [ClassicSimilarity], result of:
            0.03599117 = score(doc=209,freq=7.0), product of:
              0.061310507 = queryWeight, product of:
                1.0113716 = boost
                4.7333736 = idf(docFreq=1021, maxDocs=42740)
                0.012807176 = queryNorm
              0.587031 = fieldWeight in 209, product of:
                2.6457512 = tf(freq=7.0), with freq of:
                  7.0 = termFreq=7.0
                4.7333736 = idf(docFreq=1021, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
          0.008990611 = weight(abstract_txt:based in 209) [ClassicSimilarity], result of:
            0.008990611 = score(doc=209,freq=2.0), product of:
              0.042265255 = queryWeight, product of:
                1.028444 = boost
                3.2088501 = idf(docFreq=4693, maxDocs=42740)
                0.012807176 = queryNorm
              0.21271871 = fieldWeight in 209, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.2088501 = idf(docFreq=4693, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
          0.00915422 = weight(abstract_txt:with in 209) [ClassicSimilarity], result of:
            0.00915422 = score(doc=209,freq=5.0), product of:
              0.034690015 = queryWeight, product of:
                1.0758718 = boost
                2.5176222 = idf(docFreq=9369, maxDocs=42740)
                0.012807176 = queryNorm
              0.2638863 = fieldWeight in 209, product of:
                2.236068 = tf(freq=5.0), with freq of:
                  5.0 = termFreq=5.0
                2.5176222 = idf(docFreq=9369, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
          0.008105389 = weight(abstract_txt:using in 209) [ClassicSimilarity], result of:
            0.008105389 = score(doc=209,freq=1.0), product of:
              0.049695447 = queryWeight, product of:
                1.1151857 = boost
                3.4794931 = idf(docFreq=3580, maxDocs=42740)
                0.012807176 = queryNorm
              0.16310124 = fieldWeight in 209, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.4794931 = idf(docFreq=3580, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
          0.019735593 = weight(abstract_txt:words in 209) [ClassicSimilarity], result of:
            0.019735593 = score(doc=209,freq=1.0), product of:
              0.07857247 = queryWeight, product of:
                1.1449288 = boost
                5.358442 = idf(docFreq=546, maxDocs=42740)
                0.012807176 = queryNorm
              0.25117695 = fieldWeight in 209, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.358442 = idf(docFreq=546, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
          0.047094356 = weight(abstract_txt:frequency in 209) [ClassicSimilarity], result of:
            0.047094356 = score(doc=209,freq=3.0), product of:
              0.09728423 = queryWeight, product of:
                1.2739855 = boost
                5.962447 = idf(docFreq=298, maxDocs=42740)
                0.012807176 = queryNorm
              0.48409036 = fieldWeight in 209, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                5.962447 = idf(docFreq=298, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
          0.042690907 = weight(abstract_txt:document in 209) [ClassicSimilarity], result of:
            0.042690907 = score(doc=209,freq=8.0), product of:
              0.075219005 = queryWeight, product of:
                1.3719956 = boost
                4.280766 = idf(docFreq=1606, maxDocs=42740)
                0.012807176 = queryNorm
              0.5675548 = fieldWeight in 209, product of:
                2.828427 = tf(freq=8.0), with freq of:
                  8.0 = termFreq=8.0
                4.280766 = idf(docFreq=1606, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
          0.039979808 = weight(abstract_txt:assumption in 209) [ClassicSimilarity], result of:
            0.039979808 = score(doc=209,freq=1.0), product of:
              0.1257952 = queryWeight, product of:
                1.4486895 = boost
                6.7800884 = idf(docFreq=131, maxDocs=42740)
                0.012807176 = queryNorm
              0.31781664 = fieldWeight in 209, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.7800884 = idf(docFreq=131, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
          0.027230313 = weight(abstract_txt:performance in 209) [ClassicSimilarity], result of:
            0.027230313 = score(doc=209,freq=2.0), product of:
              0.08847606 = queryWeight, product of:
                1.4879961 = boost
                4.6426997 = idf(docFreq=1118, maxDocs=42740)
                0.012807176 = queryNorm
              0.3077704 = fieldWeight in 209, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.6426997 = idf(docFreq=1118, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
          0.03268944 = weight(abstract_txt:relevance in 209) [ClassicSimilarity], result of:
            0.03268944 = score(doc=209,freq=2.0), product of:
              0.099937625 = queryWeight, product of:
                1.5814426 = boost
                4.934262 = idf(docFreq=835, maxDocs=42740)
                0.012807176 = queryNorm
              0.32709843 = fieldWeight in 209, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.934262 = idf(docFreq=835, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
          0.018863605 = weight(abstract_txt:retrieval in 209) [ClassicSimilarity], result of:
            0.018863605 = score(doc=209,freq=2.0), product of:
              0.08212778 = queryWeight, product of:
                1.850795 = boost
                3.4648013 = idf(docFreq=3633, maxDocs=42740)
                0.012807176 = queryNorm
              0.22968605 = fieldWeight in 209, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.4648013 = idf(docFreq=3633, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
          0.026897917 = weight(abstract_txt:context in 209) [ClassicSimilarity], result of:
            0.026897917 = score(doc=209,freq=1.0), product of:
              0.13108793 = queryWeight, product of:
                2.33827 = boost
                4.377384 = idf(docFreq=1458, maxDocs=42740)
                0.012807176 = queryNorm
              0.20518988 = fieldWeight in 209, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.377384 = idf(docFreq=1458, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
          0.17195264 = weight(abstract_txt:term in 209) [ClassicSimilarity], result of:
            0.17195264 = score(doc=209,freq=7.0), product of:
              0.2871316 = queryWeight, product of:
                4.6429076 = boost
                4.8287816 = idf(docFreq=928, maxDocs=42740)
                0.012807176 = queryNorm
              0.5988635 = fieldWeight in 209, product of:
                2.6457512 = tf(freq=7.0), with freq of:
                  7.0 = termFreq=7.0
                4.8287816 = idf(docFreq=928, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
          0.1477678 = weight(abstract_txt:weights in 209) [ClassicSimilarity], result of:
            0.1477678 = score(doc=209,freq=1.0), product of:
              0.43370777 = queryWeight, product of:
                4.6591053 = boost
                7.268441 = idf(docFreq=80, maxDocs=42740)
                0.012807176 = queryNorm
              0.3407082 = fieldWeight in 209, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.268441 = idf(docFreq=80, maxDocs=42740)
                0.046875 = fieldNorm(doc=209)
        0.56 = coord(14/25)
    
  5. Witschel, H.F.: Global term weights in distributed environments (2008) 0.34
    0.34450808 = sum of:
      0.34450808 = product of:
        0.78297293 = sum of:
          0.08766445 = weight(abstract_txt:pruning in 4097) [ClassicSimilarity], result of:
            0.08766445 = score(doc=4097,freq=1.0), product of:
              0.11987909 = queryWeight, product of:
                9.360306 = idf(docFreq=9, maxDocs=42740)
                0.012807176 = queryNorm
              0.7312739 = fieldWeight in 4097, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.360306 = idf(docFreq=9, maxDocs=42740)
                0.078125 = fieldNorm(doc=4097)
          0.010595537 = weight(abstract_txt:based in 4097) [ClassicSimilarity], result of:
            0.010595537 = score(doc=4097,freq=1.0), product of:
              0.042265255 = queryWeight, product of:
                1.028444 = boost
                3.2088501 = idf(docFreq=4693, maxDocs=42740)
                0.012807176 = queryNorm
              0.2506914 = fieldWeight in 4097, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.2088501 = idf(docFreq=4693, maxDocs=42740)
                0.078125 = fieldNorm(doc=4097)
          0.0068231523 = weight(abstract_txt:with in 4097) [ClassicSimilarity], result of:
            0.0068231523 = score(doc=4097,freq=1.0), product of:
              0.034690015 = queryWeight, product of:
                1.0758718 = boost
                2.5176222 = idf(docFreq=9369, maxDocs=42740)
                0.012807176 = queryNorm
              0.19668923 = fieldWeight in 4097, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                2.5176222 = idf(docFreq=9369, maxDocs=42740)
                0.078125 = fieldNorm(doc=4097)
          0.019104585 = weight(abstract_txt:using in 4097) [ClassicSimilarity], result of:
            0.019104585 = score(doc=4097,freq=2.0), product of:
              0.049695447 = queryWeight, product of:
                1.1151857 = boost
                3.4794931 = idf(docFreq=3580, maxDocs=42740)
                0.012807176 = queryNorm
              0.3844333 = fieldWeight in 4097, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.4794931 = idf(docFreq=3580, maxDocs=42740)
                0.078125 = fieldNorm(doc=4097)
          0.032892656 = weight(abstract_txt:words in 4097) [ClassicSimilarity], result of:
            0.032892656 = score(doc=4097,freq=1.0), product of:
              0.07857247 = queryWeight, product of:
                1.1449288 = boost
                5.358442 = idf(docFreq=546, maxDocs=42740)
                0.012807176 = queryNorm
              0.41862828 = fieldWeight in 4097, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.358442 = idf(docFreq=546, maxDocs=42740)
                0.078125 = fieldNorm(doc=4097)
          0.04531657 = weight(abstract_txt:frequency in 4097) [ClassicSimilarity], result of:
            0.04531657 = score(doc=4097,freq=1.0), product of:
              0.09728423 = queryWeight, product of:
                1.2739855 = boost
                5.962447 = idf(docFreq=298, maxDocs=42740)
                0.012807176 = queryNorm
              0.4658162 = fieldWeight in 4097, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.962447 = idf(docFreq=298, maxDocs=42740)
                0.078125 = fieldNorm(doc=4097)
          0.032091234 = weight(abstract_txt:performance in 4097) [ClassicSimilarity], result of:
            0.032091234 = score(doc=4097,freq=1.0), product of:
              0.08847606 = queryWeight, product of:
                1.4879961 = boost
                4.6426997 = idf(docFreq=1118, maxDocs=42740)
                0.012807176 = queryNorm
              0.36271092 = fieldWeight in 4097, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6426997 = idf(docFreq=1118, maxDocs=42740)
                0.078125 = fieldNorm(doc=4097)
          0.10455558 = weight(abstract_txt:estimation in 4097) [ClassicSimilarity], result of:
            0.10455558 = score(doc=4097,freq=1.0), product of:
              0.16986448 = queryWeight, product of:
                1.683428 = boost
                7.8787007 = idf(docFreq=43, maxDocs=42740)
                0.012807176 = queryNorm
              0.6155235 = fieldWeight in 4097, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.8787007 = idf(docFreq=43, maxDocs=42740)
                0.078125 = fieldNorm(doc=4097)
          0.044461943 = weight(abstract_txt:retrieval in 4097) [ClassicSimilarity], result of:
            0.044461943 = score(doc=4097,freq=4.0), product of:
              0.08212778 = queryWeight, product of:
                1.850795 = boost
                3.4648013 = idf(docFreq=3633, maxDocs=42740)
                0.012807176 = queryNorm
              0.5413752 = fieldWeight in 4097, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                3.4648013 = idf(docFreq=3633, maxDocs=42740)
                0.078125 = fieldNorm(doc=4097)
          0.15318759 = weight(abstract_txt:term in 4097) [ClassicSimilarity], result of:
            0.15318759 = score(doc=4097,freq=2.0), product of:
              0.2871316 = queryWeight, product of:
                4.6429076 = boost
                4.8287816 = idf(docFreq=928, maxDocs=42740)
                0.012807176 = queryNorm
              0.53351 = fieldWeight in 4097, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.8287816 = idf(docFreq=928, maxDocs=42740)
                0.078125 = fieldNorm(doc=4097)
          0.24627964 = weight(abstract_txt:weights in 4097) [ClassicSimilarity], result of:
            0.24627964 = score(doc=4097,freq=1.0), product of:
              0.43370777 = queryWeight, product of:
                4.6591053 = boost
                7.268441 = idf(docFreq=80, maxDocs=42740)
                0.012807176 = queryNorm
              0.56784695 = fieldWeight in 4097, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.268441 = idf(docFreq=80, maxDocs=42740)
                0.078125 = fieldNorm(doc=4097)
        0.44 = coord(11/25)