Document (#38444)

Author
Mesquita, L.A.P.
Souza, R.R.
Baracho Porto, R.M.A.
Title
Noun phrases in automatic indexing: : a structural analysis of the distribution of relevant terms in doctoral theses
Source
Knowledge organization in the 21st century: between historical patterns and future prospects. Proceedings of the Thirteenth International ISKO Conference 19-22 May 2014, Kraków, Poland. Ed.: Wieslaw Babik
Imprint
Würzburg : Ergon Verlag
Year
2014
Pages
S.327-334
Series
Advances in knowledge organization; vol. 14
Abstract
The main objective of this research was to analyze whether there was a characteristic distribution behavior of relevant terms over a scientific text that could contribute as a criterion for their process of automatic indexing. The terms considered in this study were only full noun phrases contained in the texts themselves. The texts were considered a total of 98 doctoral theses of the eight areas of knowledge in a same university. Initially, 20 full noun phrases were automatically extracted from each text as candidates to be the most relevant terms, and each author of each text assigned a relevance value 0-6 (not relevant and highly relevant, respectively) for each of the 20 noun phrases sent. Only, 22.1 % of noun phrases were considered not relevant. A relevance values of the terms assigned by the authors were associated with their positions in the text. Each full noun phrases found in the text was considered as a valid linear position. The results that were obtained showed values resulting from this distribution by considering two types of position: linear, with values consolidated into ten equal consecutive parts; and structural, considering parts of the text (such as introduction, development and conclusion). As a result of considerable importance, all areas of knowledge related to the Natural Sciences showed a characteristic behavior in the distribution of relevant terms, as well as all areas of knowledge related to Social Sciences showed the same characteristic behavior of distribution, but distinct from the Natural Sciences. The difference of the distribution behavior between the Natural and Social Sciences can be clearly visualized through graphs. All behaviors, including the general behavior of all areas of knowledge together, were characterized in polynomial equations and can be applied in future as criteria for automatic indexing. Until the present date this work has become inedited of for two reasons: to present a method for characterizing the distribution of relevant terms in a scientific text, and also, through this method, pointing out a quantitative trait difference between the Natural and Social Sciences.
Content
Vgl.: http://www.ergon-verlag.de/isko_ko/downloads/aiko_vol_14_2014_45.pdf.
Theme
Automatisches Indexieren

Similar documents (author)

  1. Souza, S.d.: Informacion : utopia y realidad de la bibliotelogia (1996) 5.16
    5.1571794 = sum of:
      5.1571794 = weight(author_txt:souza in 825) [ClassicSimilarity], result of:
        5.1571794 = fieldWeight in 825, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          8.251487 = idf(docFreq=29, maxDocs=42306)
          0.625 = fieldNorm(doc=825)
    
  2. Souza, Y.d.: Reference work with international students : making the most use of the neutral question (1996) 5.16
    5.1571794 = sum of:
      5.1571794 = weight(author_txt:souza in 862) [ClassicSimilarity], result of:
        5.1571794 = fieldWeight in 862, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          8.251487 = idf(docFreq=29, maxDocs=42306)
          0.625 = fieldNorm(doc=862)
    
  3. Rocha Souza, R. = > Souza, R.R.: 5.11
    5.1053467 = sum of:
      5.1053467 = weight(author_txt:souza in 2439) [ClassicSimilarity], result of:
        5.1053467 = fieldWeight in 2439, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          8.251487 = idf(docFreq=29, maxDocs=42306)
          0.4375 = fieldNorm(doc=2439)
    
  4. Rocha Souza, R. -> Souza, R.R.: 5.11
    5.1053467 = sum of:
      5.1053467 = weight(author_txt:souza in 2880) [ClassicSimilarity], result of:
        5.1053467 = fieldWeight in 2880, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          8.251487 = idf(docFreq=29, maxDocs=42306)
          0.4375 = fieldNorm(doc=2880)
    
  5. Rocha Souza, R. => Souza, R.R.: 5.11
    5.1053467 = sum of:
      5.1053467 = weight(author_txt:souza in 538) [ClassicSimilarity], result of:
        5.1053467 = fieldWeight in 538, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          8.251487 = idf(docFreq=29, maxDocs=42306)
          0.4375 = fieldNorm(doc=538)
    

Similar documents (content)

  1. Souza, R.R.; Raghavan, K.S.: ¬A methodology for noun phrase-based automatic indexing (2006) 0.32
    0.31906658 = sum of:
      0.31906658 = product of:
        1.1395235 = sum of:
          0.026733758 = weight(abstract_txt:indexing in 1299) [ClassicSimilarity], result of:
            0.026733758 = score(doc=1299,freq=1.0), product of:
              0.07888007 = queryWeight, product of:
                1.0411505 = boost
                4.3381314 = idf(docFreq=1501, maxDocs=42306)
                0.017464297 = queryNorm
              0.3389165 = fieldWeight in 1299, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.3381314 = idf(docFreq=1501, maxDocs=42306)
                0.078125 = fieldNorm(doc=1299)
          0.020425675 = weight(abstract_txt:knowledge in 1299) [ClassicSimilarity], result of:
            0.020425675 = score(doc=1299,freq=1.0), product of:
              0.07255898 = queryWeight, product of:
                1.1530411 = boost
                3.6032572 = idf(docFreq=3131, maxDocs=42306)
                0.017464297 = queryNorm
              0.28150445 = fieldWeight in 1299, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.6032572 = idf(docFreq=3131, maxDocs=42306)
                0.078125 = fieldNorm(doc=1299)
          0.05082376 = weight(abstract_txt:text in 1299) [ClassicSimilarity], result of:
            0.05082376 = score(doc=1299,freq=1.0), product of:
              0.16055755 = queryWeight, product of:
                2.2689953 = boost
                4.0517817 = idf(docFreq=1999, maxDocs=42306)
                0.017464297 = queryNorm
              0.31654543 = fieldWeight in 1299, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.0517817 = idf(docFreq=1999, maxDocs=42306)
                0.078125 = fieldNorm(doc=1299)
          0.05112662 = weight(abstract_txt:terms in 1299) [ClassicSimilarity], result of:
            0.05112662 = score(doc=1299,freq=1.0), product of:
              0.16119476 = queryWeight, product of:
                2.2734933 = boost
                4.059814 = idf(docFreq=1983, maxDocs=42306)
                0.017464297 = queryNorm
              0.31717297 = fieldWeight in 1299, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.059814 = idf(docFreq=1983, maxDocs=42306)
                0.078125 = fieldNorm(doc=1299)
          0.0885028 = weight(abstract_txt:relevant in 1299) [ClassicSimilarity], result of:
            0.0885028 = score(doc=1299,freq=1.0), product of:
              0.24297124 = queryWeight, product of:
                2.9839506 = boost
                4.662428 = idf(docFreq=1085, maxDocs=42306)
                0.017464297 = queryNorm
              0.36425218 = fieldWeight in 1299, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.662428 = idf(docFreq=1085, maxDocs=42306)
                0.078125 = fieldNorm(doc=1299)
          0.3630831 = weight(abstract_txt:phrases in 1299) [ClassicSimilarity], result of:
            0.3630831 = score(doc=1299,freq=3.0), product of:
              0.39225417 = queryWeight, product of:
                3.2834365 = boost
                6.8405 = idf(docFreq=122, maxDocs=42306)
                0.017464297 = queryNorm
              0.92563224 = fieldWeight in 1299, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                6.8405 = idf(docFreq=122, maxDocs=42306)
                0.078125 = fieldNorm(doc=1299)
          0.5388277 = weight(abstract_txt:noun in 1299) [ClassicSimilarity], result of:
            0.5388277 = score(doc=1299,freq=3.0), product of:
              0.51034456 = queryWeight, product of:
                3.7452135 = boost
                7.8025365 = idf(docFreq=46, maxDocs=42306)
                0.017464297 = queryNorm
              1.0558116 = fieldWeight in 1299, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                7.8025365 = idf(docFreq=46, maxDocs=42306)
                0.078125 = fieldNorm(doc=1299)
        0.28 = coord(7/25)
    
  2. Kim, W.; Wilbur, W.J.: Corpus-based statistical screening for content-bearing terms (2001) 0.26
    0.25572196 = sum of:
      0.25572196 = product of:
        0.7103387 = sum of:
          0.055643164 = weight(abstract_txt:values in 189) [ClassicSimilarity], result of:
            0.055643164 = score(doc=189,freq=2.0), product of:
              0.14346886 = queryWeight, product of:
                1.4041343 = boost
                5.850566 = idf(docFreq=330, maxDocs=42306)
                0.017464297 = queryNorm
              0.3878414 = fieldWeight in 189, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.850566 = idf(docFreq=330, maxDocs=42306)
                0.046875 = fieldNorm(doc=189)
          0.033821497 = weight(abstract_txt:considered in 189) [ClassicSimilarity], result of:
            0.033821497 = score(doc=189,freq=1.0), product of:
              0.14275825 = queryWeight, product of:
                1.6173344 = boost
                5.0541754 = idf(docFreq=733, maxDocs=42306)
                0.017464297 = queryNorm
              0.23691447 = fieldWeight in 189, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.0541754 = idf(docFreq=733, maxDocs=42306)
                0.046875 = fieldNorm(doc=189)
          0.03495898 = weight(abstract_txt:natural in 189) [ClassicSimilarity], result of:
            0.03495898 = score(doc=189,freq=1.0), product of:
              0.1459414 = queryWeight, product of:
                1.6352662 = boost
                5.1102123 = idf(docFreq=693, maxDocs=42306)
                0.017464297 = queryNorm
              0.2395412 = fieldWeight in 189, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.1102123 = idf(docFreq=693, maxDocs=42306)
                0.046875 = fieldNorm(doc=189)
          0.05709878 = weight(abstract_txt:each in 189) [ClassicSimilarity], result of:
            0.05709878 = score(doc=189,freq=6.0), product of:
              0.11998956 = queryWeight, product of:
                1.6577764 = boost
                4.1444454 = idf(docFreq=1822, maxDocs=42306)
                0.017464297 = queryNorm
              0.47586456 = fieldWeight in 189, product of:
                2.4494898 = tf(freq=6.0), with freq of:
                  6.0 = termFreq=6.0
                4.1444454 = idf(docFreq=1822, maxDocs=42306)
                0.046875 = fieldNorm(doc=189)
          0.023255683 = weight(abstract_txt:were in 189) [ClassicSimilarity], result of:
            0.023255683 = score(doc=189,freq=1.0), product of:
              0.13402057 = queryWeight, product of:
                2.0730221 = boost
                3.7018294 = idf(docFreq=2837, maxDocs=42306)
                0.017464297 = queryNorm
              0.17352325 = fieldWeight in 189, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.7018294 = idf(docFreq=2837, maxDocs=42306)
                0.046875 = fieldNorm(doc=189)
          0.030494258 = weight(abstract_txt:text in 189) [ClassicSimilarity], result of:
            0.030494258 = score(doc=189,freq=1.0), product of:
              0.16055755 = queryWeight, product of:
                2.2689953 = boost
                4.0517817 = idf(docFreq=1999, maxDocs=42306)
                0.017464297 = queryNorm
              0.18992727 = fieldWeight in 189, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.0517817 = idf(docFreq=1999, maxDocs=42306)
                0.046875 = fieldNorm(doc=189)
          0.06135194 = weight(abstract_txt:terms in 189) [ClassicSimilarity], result of:
            0.06135194 = score(doc=189,freq=4.0), product of:
              0.16119476 = queryWeight, product of:
                2.2734933 = boost
                4.059814 = idf(docFreq=1983, maxDocs=42306)
                0.017464297 = queryNorm
              0.38060755 = fieldWeight in 189, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                4.059814 = idf(docFreq=1983, maxDocs=42306)
                0.046875 = fieldNorm(doc=189)
          0.08094322 = weight(abstract_txt:distribution in 189) [ClassicSimilarity], result of:
            0.08094322 = score(doc=189,freq=1.0), product of:
              0.30780265 = queryWeight, product of:
                3.1416254 = boost
                5.610051 = idf(docFreq=420, maxDocs=42306)
                0.017464297 = queryNorm
              0.26297116 = fieldWeight in 189, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.610051 = idf(docFreq=420, maxDocs=42306)
                0.046875 = fieldNorm(doc=189)
          0.33277118 = weight(abstract_txt:phrases in 189) [ClassicSimilarity], result of:
            0.33277118 = score(doc=189,freq=7.0), product of:
              0.39225417 = queryWeight, product of:
                3.2834365 = boost
                6.8405 = idf(docFreq=122, maxDocs=42306)
                0.017464297 = queryNorm
              0.848356 = fieldWeight in 189, product of:
                2.6457512 = tf(freq=7.0), with freq of:
                  7.0 = termFreq=7.0
                6.8405 = idf(docFreq=122, maxDocs=42306)
                0.046875 = fieldNorm(doc=189)
        0.36 = coord(9/25)
    
  3. Vlachidis, A.; Tudhope, D.: ¬A knowledge-based approach to information extraction for semantic interoperability in the archaeology domain (2016) 0.21
    0.2105542 = sum of:
      0.2105542 = product of:
        0.6579819 = sum of:
          0.030245794 = weight(abstract_txt:indexing in 4896) [ClassicSimilarity], result of:
            0.030245794 = score(doc=4896,freq=2.0), product of:
              0.07888007 = queryWeight, product of:
                1.0411505 = boost
                4.3381314 = idf(docFreq=1501, maxDocs=42306)
                0.017464297 = queryNorm
              0.38344026 = fieldWeight in 4896, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.3381314 = idf(docFreq=1501, maxDocs=42306)
                0.0625 = fieldNorm(doc=4896)
          0.01634054 = weight(abstract_txt:knowledge in 4896) [ClassicSimilarity], result of:
            0.01634054 = score(doc=4896,freq=1.0), product of:
              0.07255898 = queryWeight, product of:
                1.1530411 = boost
                3.6032572 = idf(docFreq=3131, maxDocs=42306)
                0.017464297 = queryNorm
              0.22520357 = fieldWeight in 4896, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.6032572 = idf(docFreq=3131, maxDocs=42306)
                0.0625 = fieldNorm(doc=4896)
          0.03674751 = weight(abstract_txt:automatic in 4896) [ClassicSimilarity], result of:
            0.03674751 = score(doc=4896,freq=1.0), product of:
              0.11315817 = queryWeight, product of:
                1.2470182 = boost
                5.1959147 = idf(docFreq=636, maxDocs=42306)
                0.017464297 = queryNorm
              0.32474467 = fieldWeight in 4896, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.1959147 = idf(docFreq=636, maxDocs=42306)
                0.0625 = fieldNorm(doc=4896)
          0.046611972 = weight(abstract_txt:natural in 4896) [ClassicSimilarity], result of:
            0.046611972 = score(doc=4896,freq=1.0), product of:
              0.1459414 = queryWeight, product of:
                1.6352662 = boost
                5.1102123 = idf(docFreq=693, maxDocs=42306)
                0.017464297 = queryNorm
              0.31938827 = fieldWeight in 4896, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.1102123 = idf(docFreq=693, maxDocs=42306)
                0.0625 = fieldNorm(doc=4896)
          0.04065901 = weight(abstract_txt:text in 4896) [ClassicSimilarity], result of:
            0.04065901 = score(doc=4896,freq=1.0), product of:
              0.16055755 = queryWeight, product of:
                2.2689953 = boost
                4.0517817 = idf(docFreq=1999, maxDocs=42306)
                0.017464297 = queryNorm
              0.25323635 = fieldWeight in 4896, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.0517817 = idf(docFreq=1999, maxDocs=42306)
                0.0625 = fieldNorm(doc=4896)
          0.07080224 = weight(abstract_txt:relevant in 4896) [ClassicSimilarity], result of:
            0.07080224 = score(doc=4896,freq=1.0), product of:
              0.24297124 = queryWeight, product of:
                2.9839506 = boost
                4.662428 = idf(docFreq=1085, maxDocs=42306)
                0.017464297 = queryNorm
              0.29140174 = fieldWeight in 4896, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.662428 = idf(docFreq=1085, maxDocs=42306)
                0.0625 = fieldNorm(doc=4896)
          0.16770092 = weight(abstract_txt:phrases in 4896) [ClassicSimilarity], result of:
            0.16770092 = score(doc=4896,freq=1.0), product of:
              0.39225417 = queryWeight, product of:
                3.2834365 = boost
                6.8405 = idf(docFreq=122, maxDocs=42306)
                0.017464297 = queryNorm
              0.42753124 = fieldWeight in 4896, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.8405 = idf(docFreq=122, maxDocs=42306)
                0.0625 = fieldNorm(doc=4896)
          0.24887387 = weight(abstract_txt:noun in 4896) [ClassicSimilarity], result of:
            0.24887387 = score(doc=4896,freq=1.0), product of:
              0.51034456 = queryWeight, product of:
                3.7452135 = boost
                7.8025365 = idf(docFreq=46, maxDocs=42306)
                0.017464297 = queryNorm
              0.48765853 = fieldWeight in 4896, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.8025365 = idf(docFreq=46, maxDocs=42306)
                0.0625 = fieldNorm(doc=4896)
        0.32 = coord(8/25)
    
  4. Salles, T.; Rocha, L.; Gonçalves, M.A.; Almeida, J.M.; Mourão, F.; Meira Jr., W.; Viegas, F.: ¬A quantitative analysis of the temporal effects on automatic text classification (2016) 0.19
    0.1900472 = sum of:
      0.1900472 = product of:
        0.52790886 = sum of:
          0.03107237 = weight(abstract_txt:full in 15) [ClassicSimilarity], result of:
            0.03107237 = score(doc=15,freq=1.0), product of:
              0.10118517 = queryWeight, product of:
                1.1792022 = boost
                4.9133477 = idf(docFreq=844, maxDocs=42306)
                0.017464297 = queryNorm
              0.30708423 = fieldWeight in 15, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.9133477 = idf(docFreq=844, maxDocs=42306)
                0.0625 = fieldNorm(doc=15)
          0.03674751 = weight(abstract_txt:automatic in 15) [ClassicSimilarity], result of:
            0.03674751 = score(doc=15,freq=1.0), product of:
              0.11315817 = queryWeight, product of:
                1.2470182 = boost
                5.1959147 = idf(docFreq=636, maxDocs=42306)
                0.017464297 = queryNorm
              0.32474467 = fieldWeight in 15, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.1959147 = idf(docFreq=636, maxDocs=42306)
                0.0625 = fieldNorm(doc=15)
          0.04509533 = weight(abstract_txt:considered in 15) [ClassicSimilarity], result of:
            0.04509533 = score(doc=15,freq=1.0), product of:
              0.14275825 = queryWeight, product of:
                1.6173344 = boost
                5.0541754 = idf(docFreq=733, maxDocs=42306)
                0.017464297 = queryNorm
              0.31588596 = fieldWeight in 15, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.0541754 = idf(docFreq=733, maxDocs=42306)
                0.0625 = fieldNorm(doc=15)
          0.043954656 = weight(abstract_txt:each in 15) [ClassicSimilarity], result of:
            0.043954656 = score(doc=15,freq=2.0), product of:
              0.11998956 = queryWeight, product of:
                1.6577764 = boost
                4.1444454 = idf(docFreq=1822, maxDocs=42306)
                0.017464297 = queryNorm
              0.36632067 = fieldWeight in 15, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.1444454 = idf(docFreq=1822, maxDocs=42306)
                0.0625 = fieldNorm(doc=15)
          0.06604848 = weight(abstract_txt:behavior in 15) [ClassicSimilarity], result of:
            0.06604848 = score(doc=15,freq=1.0), product of:
              0.19833168 = queryWeight, product of:
                2.1313279 = boost
                5.3283253 = idf(docFreq=557, maxDocs=42306)
                0.017464297 = queryNorm
              0.33302033 = fieldWeight in 15, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.3283253 = idf(docFreq=557, maxDocs=42306)
                0.0625 = fieldNorm(doc=15)
          0.04065901 = weight(abstract_txt:text in 15) [ClassicSimilarity], result of:
            0.04065901 = score(doc=15,freq=1.0), product of:
              0.16055755 = queryWeight, product of:
                2.2689953 = boost
                4.0517817 = idf(docFreq=1999, maxDocs=42306)
                0.017464297 = queryNorm
              0.25323635 = fieldWeight in 15, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.0517817 = idf(docFreq=1999, maxDocs=42306)
                0.0625 = fieldNorm(doc=15)
          0.040901296 = weight(abstract_txt:terms in 15) [ClassicSimilarity], result of:
            0.040901296 = score(doc=15,freq=1.0), product of:
              0.16119476 = queryWeight, product of:
                2.2734933 = boost
                4.059814 = idf(docFreq=1983, maxDocs=42306)
                0.017464297 = queryNorm
              0.25373837 = fieldWeight in 15, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.059814 = idf(docFreq=1983, maxDocs=42306)
                0.0625 = fieldNorm(doc=15)
          0.07080224 = weight(abstract_txt:relevant in 15) [ClassicSimilarity], result of:
            0.07080224 = score(doc=15,freq=1.0), product of:
              0.24297124 = queryWeight, product of:
                2.9839506 = boost
                4.662428 = idf(docFreq=1085, maxDocs=42306)
                0.017464297 = queryNorm
              0.29140174 = fieldWeight in 15, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.662428 = idf(docFreq=1085, maxDocs=42306)
                0.0625 = fieldNorm(doc=15)
          0.15262799 = weight(abstract_txt:distribution in 15) [ClassicSimilarity], result of:
            0.15262799 = score(doc=15,freq=2.0), product of:
              0.30780265 = queryWeight, product of:
                3.1416254 = boost
                5.610051 = idf(docFreq=420, maxDocs=42306)
                0.017464297 = queryNorm
              0.49586314 = fieldWeight in 15, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.610051 = idf(docFreq=420, maxDocs=42306)
                0.0625 = fieldNorm(doc=15)
        0.36 = coord(9/25)
    
  5. Spitkovsky, V.; Norvig, P.: From words to concepts and back : dictionaries for linking text, entities and ideas (2012) 0.18
    0.1847469 = sum of:
      0.1847469 = product of:
        0.5773341 = sum of:
          0.030544046 = weight(abstract_txt:areas in 2338) [ClassicSimilarity], result of:
            0.030544046 = score(doc=2338,freq=1.0), product of:
              0.1333799 = queryWeight, product of:
                1.5633075 = boost
                4.885341 = idf(docFreq=868, maxDocs=42306)
                0.017464297 = queryNorm
              0.22900036 = fieldWeight in 2338, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.885341 = idf(docFreq=868, maxDocs=42306)
                0.046875 = fieldNorm(doc=2338)
          0.04943946 = weight(abstract_txt:natural in 2338) [ClassicSimilarity], result of:
            0.04943946 = score(doc=2338,freq=2.0), product of:
              0.1459414 = queryWeight, product of:
                1.6352662 = boost
                5.1102123 = idf(docFreq=693, maxDocs=42306)
                0.017464297 = queryNorm
              0.3387624 = fieldWeight in 2338, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.1102123 = idf(docFreq=693, maxDocs=42306)
                0.046875 = fieldNorm(doc=2338)
          0.04037493 = weight(abstract_txt:each in 2338) [ClassicSimilarity], result of:
            0.04037493 = score(doc=2338,freq=3.0), product of:
              0.11998956 = queryWeight, product of:
                1.6577764 = boost
                4.1444454 = idf(docFreq=1822, maxDocs=42306)
                0.017464297 = queryNorm
              0.33648703 = fieldWeight in 2338, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.1444454 = idf(docFreq=1822, maxDocs=42306)
                0.046875 = fieldNorm(doc=2338)
          0.023255683 = weight(abstract_txt:were in 2338) [ClassicSimilarity], result of:
            0.023255683 = score(doc=2338,freq=1.0), product of:
              0.13402057 = queryWeight, product of:
                2.0730221 = boost
                3.7018294 = idf(docFreq=2837, maxDocs=42306)
                0.017464297 = queryNorm
              0.17352325 = fieldWeight in 2338, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.7018294 = idf(docFreq=2837, maxDocs=42306)
                0.046875 = fieldNorm(doc=2338)
          0.06818724 = weight(abstract_txt:text in 2338) [ClassicSimilarity], result of:
            0.06818724 = score(doc=2338,freq=5.0), product of:
              0.16055755 = queryWeight, product of:
                2.2689953 = boost
                4.0517817 = idf(docFreq=1999, maxDocs=42306)
                0.017464297 = queryNorm
              0.4246903 = fieldWeight in 2338, product of:
                2.236068 = tf(freq=5.0), with freq of:
                  5.0 = termFreq=5.0
                4.0517817 = idf(docFreq=1999, maxDocs=42306)
                0.046875 = fieldNorm(doc=2338)
          0.05310168 = weight(abstract_txt:relevant in 2338) [ClassicSimilarity], result of:
            0.05310168 = score(doc=2338,freq=1.0), product of:
              0.24297124 = queryWeight, product of:
                2.9839506 = boost
                4.662428 = idf(docFreq=1085, maxDocs=42306)
                0.017464297 = queryNorm
              0.21855131 = fieldWeight in 2338, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.662428 = idf(docFreq=1085, maxDocs=42306)
                0.046875 = fieldNorm(doc=2338)
          0.12577568 = weight(abstract_txt:phrases in 2338) [ClassicSimilarity], result of:
            0.12577568 = score(doc=2338,freq=1.0), product of:
              0.39225417 = queryWeight, product of:
                3.2834365 = boost
                6.8405 = idf(docFreq=122, maxDocs=42306)
                0.017464297 = queryNorm
              0.32064843 = fieldWeight in 2338, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.8405 = idf(docFreq=122, maxDocs=42306)
                0.046875 = fieldNorm(doc=2338)
          0.18665542 = weight(abstract_txt:noun in 2338) [ClassicSimilarity], result of:
            0.18665542 = score(doc=2338,freq=1.0), product of:
              0.51034456 = queryWeight, product of:
                3.7452135 = boost
                7.8025365 = idf(docFreq=46, maxDocs=42306)
                0.017464297 = queryNorm
              0.3657439 = fieldWeight in 2338, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.8025365 = idf(docFreq=46, maxDocs=42306)
                0.046875 = fieldNorm(doc=2338)
        0.32 = coord(8/25)