Document (#38444)

Author
Mesquita, L.A.P.
Souza, R.R.
Baracho Porto, R.M.A.
Title
Noun phrases in automatic indexing: : a structural analysis of the distribution of relevant terms in doctoral theses
Source
Knowledge organization in the 21st century: between historical patterns and future prospects. Proceedings of the Thirteenth International ISKO Conference 19-22 May 2014, Kraków, Poland. Ed.: Wieslaw Babik
Imprint
Würzburg : Ergon Verlag
Year
2014
Pages
S.327-334
Series
Advances in knowledge organization; vol. 14
Abstract
The main objective of this research was to analyze whether there was a characteristic distribution behavior of relevant terms over a scientific text that could contribute as a criterion for their process of automatic indexing. The terms considered in this study were only full noun phrases contained in the texts themselves. The texts were considered a total of 98 doctoral theses of the eight areas of knowledge in a same university. Initially, 20 full noun phrases were automatically extracted from each text as candidates to be the most relevant terms, and each author of each text assigned a relevance value 0-6 (not relevant and highly relevant, respectively) for each of the 20 noun phrases sent. Only, 22.1 % of noun phrases were considered not relevant. A relevance values of the terms assigned by the authors were associated with their positions in the text. Each full noun phrases found in the text was considered as a valid linear position. The results that were obtained showed values resulting from this distribution by considering two types of position: linear, with values consolidated into ten equal consecutive parts; and structural, considering parts of the text (such as introduction, development and conclusion). As a result of considerable importance, all areas of knowledge related to the Natural Sciences showed a characteristic behavior in the distribution of relevant terms, as well as all areas of knowledge related to Social Sciences showed the same characteristic behavior of distribution, but distinct from the Natural Sciences. The difference of the distribution behavior between the Natural and Social Sciences can be clearly visualized through graphs. All behaviors, including the general behavior of all areas of knowledge together, were characterized in polynomial equations and can be applied in future as criteria for automatic indexing. Until the present date this work has become inedited of for two reasons: to present a method for characterizing the distribution of relevant terms in a scientific text, and also, through this method, pointing out a quantitative trait difference between the Natural and Social Sciences.
Content
Vgl.: http://www.ergon-verlag.de/isko_ko/downloads/aiko_vol_14_2014_45.pdf.
Theme
Automatisches Indexieren

Similar documents (author)

  1. Barcellos Almeida, M.; Souza, R.R.; Porto, R.B.: Looking for the identity of information science in the age of big data, computing clouds and social networks (2015) 4.73
    4.7339735 = sum of:
      4.7339735 = sum of:
        1.7402284 = weight(author_txt:souza in 4918) [ClassicSimilarity], result of:
          1.7402284 = score(doc=4918,freq=1.0), product of:
            0.571539 = queryWeight, product of:
              8.119497 = idf(docFreq=34, maxDocs=43254)
              0.07039093 = queryNorm
            3.0448115 = fieldWeight in 4918, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              8.119497 = idf(docFreq=34, maxDocs=43254)
              0.375 = fieldNorm(doc=4918)
        2.9937449 = weight(author_txt:porto in 4918) [ClassicSimilarity], result of:
          2.9937449 = score(doc=4918,freq=1.0), product of:
            0.8205749 = queryWeight, product of:
              1.198219 = boost
              9.728935 = idf(docFreq=6, maxDocs=43254)
              0.07039093 = queryNorm
            3.6483507 = fieldWeight in 4918, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              9.728935 = idf(docFreq=6, maxDocs=43254)
              0.375 = fieldNorm(doc=4918)
    
  2. Porto, R. => Porto, R.B.: 2.82
    2.8225298 = sum of:
      2.8225298 = product of:
        5.6450596 = sum of:
          5.6450596 = weight(author_txt:porto in 4851) [ClassicSimilarity], result of:
            5.6450596 = score(doc=4851,freq=2.0), product of:
              0.8205749 = queryWeight, product of:
                1.198219 = boost
                9.728935 = idf(docFreq=6, maxDocs=43254)
                0.07039093 = queryNorm
              6.879396 = fieldWeight in 4851, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                9.728935 = idf(docFreq=6, maxDocs=43254)
                0.5 = fieldNorm(doc=4851)
        0.5 = coord(1/2)
    
  3. Gomez, I. Porto- => Porto-Gomez, I.: 2.12
    2.1168973 = sum of:
      2.1168973 = product of:
        4.2337947 = sum of:
          4.2337947 = weight(author_txt:porto in 373) [ClassicSimilarity], result of:
            4.2337947 = score(doc=373,freq=2.0), product of:
              0.8205749 = queryWeight, product of:
                1.198219 = boost
                9.728935 = idf(docFreq=6, maxDocs=43254)
                0.07039093 = queryNorm
              5.159547 = fieldWeight in 373, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                9.728935 = idf(docFreq=6, maxDocs=43254)
                0.375 = fieldNorm(doc=373)
        0.5 = coord(1/2)
    
  4. Dal Porto, S.; Marchitelli, A.: ¬The functionality and flexibility of traditional classification schemes applied to a Content Management System (CMS) : facets, DDC, JITA (2006) 1.75
    1.7463512 = sum of:
      1.7463512 = product of:
        3.4927025 = sum of:
          3.4927025 = weight(author_txt:porto in 1300) [ClassicSimilarity], result of:
            3.4927025 = score(doc=1300,freq=1.0), product of:
              0.8205749 = queryWeight, product of:
                1.198219 = boost
                9.728935 = idf(docFreq=6, maxDocs=43254)
                0.07039093 = queryNorm
              4.256409 = fieldWeight in 1300, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.728935 = idf(docFreq=6, maxDocs=43254)
                0.4375 = fieldNorm(doc=1300)
        0.5 = coord(1/2)
    
  5. Souza, S.d.: Informacion : utopia y realidad de la bibliotelogia (1996) 1.45
    1.4501904 = sum of:
      1.4501904 = product of:
        2.9003808 = sum of:
          2.9003808 = weight(author_txt:souza in 1825) [ClassicSimilarity], result of:
            2.9003808 = score(doc=1825,freq=1.0), product of:
              0.571539 = queryWeight, product of:
                8.119497 = idf(docFreq=34, maxDocs=43254)
                0.07039093 = queryNorm
              5.074686 = fieldWeight in 1825, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.119497 = idf(docFreq=34, maxDocs=43254)
                0.625 = fieldNorm(doc=1825)
        0.5 = coord(1/2)
    

Similar documents (content)

  1. Souza, R.R.; Raghavan, K.S.: ¬A methodology for noun phrase-based automatic indexing (2006) 0.32
    0.31931928 = sum of:
      0.31931928 = product of:
        1.140426 = sum of:
          0.027024042 = weight(abstract_txt:indexing in 1299) [ClassicSimilarity], result of:
            0.027024042 = score(doc=1299,freq=1.0), product of:
              0.07960878 = queryWeight, product of:
                1.0472118 = boost
                4.345095 = idf(docFreq=1524, maxDocs=43254)
                0.017495533 = queryNorm
              0.33946055 = fieldWeight in 1299, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.345095 = idf(docFreq=1524, maxDocs=43254)
                0.078125 = fieldNorm(doc=1299)
          0.020145446 = weight(abstract_txt:knowledge in 1299) [ClassicSimilarity], result of:
            0.020145446 = score(doc=1299,freq=1.0), product of:
              0.072037436 = queryWeight, product of:
                1.1502771 = boost
                3.5795512 = idf(docFreq=3278, maxDocs=43254)
                0.017495533 = queryNorm
              0.27965245 = fieldWeight in 1299, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.5795512 = idf(docFreq=3278, maxDocs=43254)
                0.078125 = fieldNorm(doc=1299)
          0.05105167 = weight(abstract_txt:text in 1299) [ClassicSimilarity], result of:
            0.05105167 = score(doc=1299,freq=1.0), product of:
              0.16135892 = queryWeight, product of:
                2.2773979 = boost
                4.049738 = idf(docFreq=2048, maxDocs=43254)
                0.017495533 = queryNorm
              0.31638578 = fieldWeight in 1299, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.049738 = idf(docFreq=2048, maxDocs=43254)
                0.078125 = fieldNorm(doc=1299)
          0.0513674 = weight(abstract_txt:terms in 1299) [ClassicSimilarity], result of:
            0.0513674 = score(doc=1299,freq=1.0), product of:
              0.16202353 = queryWeight, product of:
                2.282083 = boost
                4.058069 = idf(docFreq=2031, maxDocs=43254)
                0.017495533 = queryNorm
              0.31703666 = fieldWeight in 1299, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.058069 = idf(docFreq=2031, maxDocs=43254)
                0.078125 = fieldNorm(doc=1299)
          0.08818361 = weight(abstract_txt:relevant in 1299) [ClassicSimilarity], result of:
            0.08818361 = score(doc=1299,freq=1.0), product of:
              0.24287097 = queryWeight, product of:
                2.9869378 = boost
                4.6475306 = idf(docFreq=1126, maxDocs=43254)
                0.017495533 = queryNorm
              0.3630883 = fieldWeight in 1299, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6475306 = idf(docFreq=1126, maxDocs=43254)
                0.078125 = fieldNorm(doc=1299)
          0.36882517 = weight(abstract_txt:phrases in 1299) [ClassicSimilarity], result of:
            0.36882517 = score(doc=1299,freq=3.0), product of:
              0.39717087 = queryWeight, product of:
                3.3079407 = boost
                6.8626604 = idf(docFreq=122, maxDocs=43254)
                0.017495533 = queryNorm
              0.92863095 = fieldWeight in 1299, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                6.8626604 = idf(docFreq=122, maxDocs=43254)
                0.078125 = fieldNorm(doc=1299)
          0.53382874 = weight(abstract_txt:noun in 1299) [ClassicSimilarity], result of:
            0.53382874 = score(doc=1299,freq=3.0), product of:
              0.5081965 = queryWeight, product of:
                3.7418368 = boost
                7.762822 = idf(docFreq=49, maxDocs=43254)
                0.017495533 = queryNorm
              1.0504377 = fieldWeight in 1299, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                7.762822 = idf(docFreq=49, maxDocs=43254)
                0.078125 = fieldNorm(doc=1299)
        0.28 = coord(7/25)
    
  2. Kim, W.; Wilbur, W.J.: Corpus-based statistical screening for content-bearing terms (2001) 0.26
    0.2577091 = sum of:
      0.2577091 = product of:
        0.7158586 = sum of:
          0.05575939 = weight(abstract_txt:values in 189) [ClassicSimilarity], result of:
            0.05575939 = score(doc=189,freq=2.0), product of:
              0.14395562 = queryWeight, product of:
                1.4082129 = boost
                5.8429627 = idf(docFreq=340, maxDocs=43254)
                0.017495533 = queryNorm
              0.38733736 = fieldWeight in 189, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.8429627 = idf(docFreq=340, maxDocs=43254)
                0.046875 = fieldNorm(doc=189)
          0.033638213 = weight(abstract_txt:considered in 189) [ClassicSimilarity], result of:
            0.033638213 = score(doc=189,freq=1.0), product of:
              0.14252622 = queryWeight, product of:
                1.6179711 = boost
                5.0349693 = idf(docFreq=764, maxDocs=43254)
                0.017495533 = queryNorm
              0.23601419 = fieldWeight in 189, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.0349693 = idf(docFreq=764, maxDocs=43254)
                0.046875 = fieldNorm(doc=189)
          0.03509785 = weight(abstract_txt:natural in 189) [ClassicSimilarity], result of:
            0.03509785 = score(doc=189,freq=1.0), product of:
              0.14661999 = queryWeight, product of:
                1.6410431 = boost
                5.106767 = idf(docFreq=711, maxDocs=43254)
                0.017495533 = queryNorm
              0.2393797 = fieldWeight in 189, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.106767 = idf(docFreq=711, maxDocs=43254)
                0.046875 = fieldNorm(doc=189)
          0.05664668 = weight(abstract_txt:each in 189) [ClassicSimilarity], result of:
            0.05664668 = score(doc=189,freq=6.0), product of:
              0.11959382 = queryWeight, product of:
                1.657039 = boost
                4.125236 = idf(docFreq=1899, maxDocs=43254)
                0.017495533 = queryNorm
              0.47365892 = fieldWeight in 189, product of:
                2.4494898 = tf(freq=6.0), with freq of:
                  6.0 = termFreq=6.0
                4.125236 = idf(docFreq=1899, maxDocs=43254)
                0.046875 = fieldNorm(doc=189)
          0.023038562 = weight(abstract_txt:were in 189) [ClassicSimilarity], result of:
            0.023038562 = score(doc=189,freq=1.0), product of:
              0.1334512 = queryWeight, product of:
                2.0711122 = boost
                3.6829145 = idf(docFreq=2956, maxDocs=43254)
                0.017495533 = queryNorm
              0.17263661 = fieldWeight in 189, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.6829145 = idf(docFreq=2956, maxDocs=43254)
                0.046875 = fieldNorm(doc=189)
          0.030631 = weight(abstract_txt:text in 189) [ClassicSimilarity], result of:
            0.030631 = score(doc=189,freq=1.0), product of:
              0.16135892 = queryWeight, product of:
                2.2773979 = boost
                4.049738 = idf(docFreq=2048, maxDocs=43254)
                0.017495533 = queryNorm
              0.18983147 = fieldWeight in 189, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.049738 = idf(docFreq=2048, maxDocs=43254)
                0.046875 = fieldNorm(doc=189)
          0.061640877 = weight(abstract_txt:terms in 189) [ClassicSimilarity], result of:
            0.061640877 = score(doc=189,freq=4.0), product of:
              0.16202353 = queryWeight, product of:
                2.282083 = boost
                4.058069 = idf(docFreq=2031, maxDocs=43254)
                0.017495533 = queryNorm
              0.380444 = fieldWeight in 189, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                4.058069 = idf(docFreq=2031, maxDocs=43254)
                0.046875 = fieldNorm(doc=189)
          0.08137212 = weight(abstract_txt:distribution in 189) [ClassicSimilarity], result of:
            0.08137212 = score(doc=189,freq=1.0), product of:
              0.30950615 = queryWeight, product of:
                3.1541114 = boost
                5.608737 = idf(docFreq=430, maxDocs=43254)
                0.017495533 = queryNorm
              0.26290953 = fieldWeight in 189, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.608737 = idf(docFreq=430, maxDocs=43254)
                0.046875 = fieldNorm(doc=189)
          0.33803385 = weight(abstract_txt:phrases in 189) [ClassicSimilarity], result of:
            0.33803385 = score(doc=189,freq=7.0), product of:
              0.39717087 = queryWeight, product of:
                3.3079407 = boost
                6.8626604 = idf(docFreq=122, maxDocs=43254)
                0.017495533 = queryNorm
              0.8511044 = fieldWeight in 189, product of:
                2.6457512 = tf(freq=7.0), with freq of:
                  7.0 = termFreq=7.0
                6.8626604 = idf(docFreq=122, maxDocs=43254)
                0.046875 = fieldNorm(doc=189)
        0.36 = coord(9/25)
    
  3. Vlachidis, A.; Tudhope, D.: ¬A knowledge-based approach to information extraction for semantic interoperability in the archaeology domain (2016) 0.21
    0.21080674 = sum of:
      0.21080674 = product of:
        0.6587711 = sum of:
          0.030574212 = weight(abstract_txt:indexing in 4360) [ClassicSimilarity], result of:
            0.030574212 = score(doc=4360,freq=2.0), product of:
              0.07960878 = queryWeight, product of:
                1.0472118 = boost
                4.345095 = idf(docFreq=1524, maxDocs=43254)
                0.017495533 = queryNorm
              0.38405576 = fieldWeight in 4360, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.345095 = idf(docFreq=1524, maxDocs=43254)
                0.0625 = fieldNorm(doc=4360)
          0.016116356 = weight(abstract_txt:knowledge in 4360) [ClassicSimilarity], result of:
            0.016116356 = score(doc=4360,freq=1.0), product of:
              0.072037436 = queryWeight, product of:
                1.1502771 = boost
                3.5795512 = idf(docFreq=3278, maxDocs=43254)
                0.017495533 = queryNorm
              0.22372195 = fieldWeight in 4360, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.5795512 = idf(docFreq=3278, maxDocs=43254)
                0.0625 = fieldNorm(doc=4360)
          0.03697719 = weight(abstract_txt:automatic in 4360) [ClassicSimilarity], result of:
            0.03697719 = score(doc=4360,freq=1.0), product of:
              0.11385621 = queryWeight, product of:
                1.2523692 = boost
                5.1963353 = idf(docFreq=650, maxDocs=43254)
                0.017495533 = queryNorm
              0.32477096 = fieldWeight in 4360, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.1963353 = idf(docFreq=650, maxDocs=43254)
                0.0625 = fieldNorm(doc=4360)
          0.046797134 = weight(abstract_txt:natural in 4360) [ClassicSimilarity], result of:
            0.046797134 = score(doc=4360,freq=1.0), product of:
              0.14661999 = queryWeight, product of:
                1.6410431 = boost
                5.106767 = idf(docFreq=711, maxDocs=43254)
                0.017495533 = queryNorm
              0.31917295 = fieldWeight in 4360, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.106767 = idf(docFreq=711, maxDocs=43254)
                0.0625 = fieldNorm(doc=4360)
          0.040841334 = weight(abstract_txt:text in 4360) [ClassicSimilarity], result of:
            0.040841334 = score(doc=4360,freq=1.0), product of:
              0.16135892 = queryWeight, product of:
                2.2773979 = boost
                4.049738 = idf(docFreq=2048, maxDocs=43254)
                0.017495533 = queryNorm
              0.25310862 = fieldWeight in 4360, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.049738 = idf(docFreq=2048, maxDocs=43254)
                0.0625 = fieldNorm(doc=4360)
          0.07054689 = weight(abstract_txt:relevant in 4360) [ClassicSimilarity], result of:
            0.07054689 = score(doc=4360,freq=1.0), product of:
              0.24287097 = queryWeight, product of:
                2.9869378 = boost
                4.6475306 = idf(docFreq=1126, maxDocs=43254)
                0.017495533 = queryNorm
              0.29047066 = fieldWeight in 4360, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6475306 = idf(docFreq=1126, maxDocs=43254)
                0.0625 = fieldNorm(doc=4360)
          0.17035306 = weight(abstract_txt:phrases in 4360) [ClassicSimilarity], result of:
            0.17035306 = score(doc=4360,freq=1.0), product of:
              0.39717087 = queryWeight, product of:
                3.3079407 = boost
                6.8626604 = idf(docFreq=122, maxDocs=43254)
                0.017495533 = queryNorm
              0.42891628 = fieldWeight in 4360, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.8626604 = idf(docFreq=122, maxDocs=43254)
                0.0625 = fieldNorm(doc=4360)
          0.24656492 = weight(abstract_txt:noun in 4360) [ClassicSimilarity], result of:
            0.24656492 = score(doc=4360,freq=1.0), product of:
              0.5081965 = queryWeight, product of:
                3.7418368 = boost
                7.762822 = idf(docFreq=49, maxDocs=43254)
                0.017495533 = queryNorm
              0.48517638 = fieldWeight in 4360, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.762822 = idf(docFreq=49, maxDocs=43254)
                0.0625 = fieldNorm(doc=4360)
        0.32 = coord(8/25)
    
  4. Salles, T.; Rocha, L.; Gonçalves, M.A.; Almeida, J.M.; Mourão, F.; Meira Jr., W.; Viegas, F.: ¬A quantitative analysis of the temporal effects on automatic text classification (2016) 0.19
    0.18957247 = sum of:
      0.18957247 = product of:
        0.52659017 = sum of:
          0.031323977 = weight(abstract_txt:full in 4479) [ClassicSimilarity], result of:
            0.031323977 = score(doc=4479,freq=1.0), product of:
              0.10193391 = queryWeight, product of:
                1.1849864 = boost
                4.9167504 = idf(docFreq=860, maxDocs=43254)
                0.017495533 = queryNorm
              0.3072969 = fieldWeight in 4479, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.9167504 = idf(docFreq=860, maxDocs=43254)
                0.0625 = fieldNorm(doc=4479)
          0.03697719 = weight(abstract_txt:automatic in 4479) [ClassicSimilarity], result of:
            0.03697719 = score(doc=4479,freq=1.0), product of:
              0.11385621 = queryWeight, product of:
                1.2523692 = boost
                5.1963353 = idf(docFreq=650, maxDocs=43254)
                0.017495533 = queryNorm
              0.32477096 = fieldWeight in 4479, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.1963353 = idf(docFreq=650, maxDocs=43254)
                0.0625 = fieldNorm(doc=4479)
          0.04485095 = weight(abstract_txt:considered in 4479) [ClassicSimilarity], result of:
            0.04485095 = score(doc=4479,freq=1.0), product of:
              0.14252622 = queryWeight, product of:
                1.6179711 = boost
                5.0349693 = idf(docFreq=764, maxDocs=43254)
                0.017495533 = queryNorm
              0.31468558 = fieldWeight in 4479, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.0349693 = idf(docFreq=764, maxDocs=43254)
                0.0625 = fieldNorm(doc=4479)
          0.043606635 = weight(abstract_txt:each in 4479) [ClassicSimilarity], result of:
            0.043606635 = score(doc=4479,freq=2.0), product of:
              0.11959382 = queryWeight, product of:
                1.657039 = boost
                4.125236 = idf(docFreq=1899, maxDocs=43254)
                0.017495533 = queryNorm
              0.3646228 = fieldWeight in 4479, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.125236 = idf(docFreq=1899, maxDocs=43254)
                0.0625 = fieldNorm(doc=4479)
          0.06391252 = weight(abstract_txt:behavior in 4479) [ClassicSimilarity], result of:
            0.06391252 = score(doc=4479,freq=1.0), product of:
              0.19442002 = queryWeight, product of:
                2.1127536 = boost
                5.259748 = idf(docFreq=610, maxDocs=43254)
                0.017495533 = queryNorm
              0.32873425 = fieldWeight in 4479, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.259748 = idf(docFreq=610, maxDocs=43254)
                0.0625 = fieldNorm(doc=4479)
          0.040841334 = weight(abstract_txt:text in 4479) [ClassicSimilarity], result of:
            0.040841334 = score(doc=4479,freq=1.0), product of:
              0.16135892 = queryWeight, product of:
                2.2773979 = boost
                4.049738 = idf(docFreq=2048, maxDocs=43254)
                0.017495533 = queryNorm
              0.25310862 = fieldWeight in 4479, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.049738 = idf(docFreq=2048, maxDocs=43254)
                0.0625 = fieldNorm(doc=4479)
          0.04109392 = weight(abstract_txt:terms in 4479) [ClassicSimilarity], result of:
            0.04109392 = score(doc=4479,freq=1.0), product of:
              0.16202353 = queryWeight, product of:
                2.282083 = boost
                4.058069 = idf(docFreq=2031, maxDocs=43254)
                0.017495533 = queryNorm
              0.25362933 = fieldWeight in 4479, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.058069 = idf(docFreq=2031, maxDocs=43254)
                0.0625 = fieldNorm(doc=4479)
          0.07054689 = weight(abstract_txt:relevant in 4479) [ClassicSimilarity], result of:
            0.07054689 = score(doc=4479,freq=1.0), product of:
              0.24287097 = queryWeight, product of:
                2.9869378 = boost
                4.6475306 = idf(docFreq=1126, maxDocs=43254)
                0.017495533 = queryNorm
              0.29047066 = fieldWeight in 4479, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6475306 = idf(docFreq=1126, maxDocs=43254)
                0.0625 = fieldNorm(doc=4479)
          0.15343675 = weight(abstract_txt:distribution in 4479) [ClassicSimilarity], result of:
            0.15343675 = score(doc=4479,freq=2.0), product of:
              0.30950615 = queryWeight, product of:
                3.1541114 = boost
                5.608737 = idf(docFreq=430, maxDocs=43254)
                0.017495533 = queryNorm
              0.495747 = fieldWeight in 4479, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.608737 = idf(docFreq=430, maxDocs=43254)
                0.0625 = fieldNorm(doc=4479)
        0.36 = coord(9/25)
    
  5. Spitkovsky, V.; Norvig, P.: From words to concepts and back : dictionaries for linking text, entities and ideas (2012) 0.18
    0.18475135 = sum of:
      0.18475135 = product of:
        0.577348 = sum of:
          0.030526662 = weight(abstract_txt:areas in 1802) [ClassicSimilarity], result of:
            0.030526662 = score(doc=1802,freq=1.0), product of:
              0.13359568 = queryWeight, product of:
                1.566461 = boost
                4.874675 = idf(docFreq=897, maxDocs=43254)
                0.017495533 = queryNorm
              0.22850038 = fieldWeight in 1802, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.874675 = idf(docFreq=897, maxDocs=43254)
                0.046875 = fieldNorm(doc=1802)
          0.049635857 = weight(abstract_txt:natural in 1802) [ClassicSimilarity], result of:
            0.049635857 = score(doc=1802,freq=2.0), product of:
              0.14661999 = queryWeight, product of:
                1.6410431 = boost
                5.106767 = idf(docFreq=711, maxDocs=43254)
                0.017495533 = queryNorm
              0.33853403 = fieldWeight in 1802, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.106767 = idf(docFreq=711, maxDocs=43254)
                0.046875 = fieldNorm(doc=1802)
          0.04005525 = weight(abstract_txt:each in 1802) [ClassicSimilarity], result of:
            0.04005525 = score(doc=1802,freq=3.0), product of:
              0.11959382 = queryWeight, product of:
                1.657039 = boost
                4.125236 = idf(docFreq=1899, maxDocs=43254)
                0.017495533 = queryNorm
              0.3349274 = fieldWeight in 1802, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.125236 = idf(docFreq=1899, maxDocs=43254)
                0.046875 = fieldNorm(doc=1802)
          0.023038562 = weight(abstract_txt:were in 1802) [ClassicSimilarity], result of:
            0.023038562 = score(doc=1802,freq=1.0), product of:
              0.1334512 = queryWeight, product of:
                2.0711122 = boost
                3.6829145 = idf(docFreq=2956, maxDocs=43254)
                0.017495533 = queryNorm
              0.17263661 = fieldWeight in 1802, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.6829145 = idf(docFreq=2956, maxDocs=43254)
                0.046875 = fieldNorm(doc=1802)
          0.068493 = weight(abstract_txt:text in 1802) [ClassicSimilarity], result of:
            0.068493 = score(doc=1802,freq=5.0), product of:
              0.16135892 = queryWeight, product of:
                2.2773979 = boost
                4.049738 = idf(docFreq=2048, maxDocs=43254)
                0.017495533 = queryNorm
              0.4244761 = fieldWeight in 1802, product of:
                2.236068 = tf(freq=5.0), with freq of:
                  5.0 = termFreq=5.0
                4.049738 = idf(docFreq=2048, maxDocs=43254)
                0.046875 = fieldNorm(doc=1802)
          0.052910168 = weight(abstract_txt:relevant in 1802) [ClassicSimilarity], result of:
            0.052910168 = score(doc=1802,freq=1.0), product of:
              0.24287097 = queryWeight, product of:
                2.9869378 = boost
                4.6475306 = idf(docFreq=1126, maxDocs=43254)
                0.017495533 = queryNorm
              0.217853 = fieldWeight in 1802, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6475306 = idf(docFreq=1126, maxDocs=43254)
                0.046875 = fieldNorm(doc=1802)
          0.12776479 = weight(abstract_txt:phrases in 1802) [ClassicSimilarity], result of:
            0.12776479 = score(doc=1802,freq=1.0), product of:
              0.39717087 = queryWeight, product of:
                3.3079407 = boost
                6.8626604 = idf(docFreq=122, maxDocs=43254)
                0.017495533 = queryNorm
              0.32168722 = fieldWeight in 1802, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.8626604 = idf(docFreq=122, maxDocs=43254)
                0.046875 = fieldNorm(doc=1802)
          0.18492371 = weight(abstract_txt:noun in 1802) [ClassicSimilarity], result of:
            0.18492371 = score(doc=1802,freq=1.0), product of:
              0.5081965 = queryWeight, product of:
                3.7418368 = boost
                7.762822 = idf(docFreq=49, maxDocs=43254)
                0.017495533 = queryNorm
              0.3638823 = fieldWeight in 1802, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.762822 = idf(docFreq=49, maxDocs=43254)
                0.046875 = fieldNorm(doc=1802)
        0.32 = coord(8/25)