Document (#32830)

Author
Peng, F.
Huang, X.
Title
Machine learning for Asian language text classification
Source
Journal of documentation. 63(2007) no.3, S.378-397
Year
2007
Abstract
Purpose - The purpose of this research is to compare several machine learning techniques on the task of Asian language text classification, such as Chinese and Japanese where no word boundary information is available in written text. The paper advocates a simple language modeling based approach for this task. Design/methodology/approach - Naïve Bayes, maximum entropy model, support vector machines, and language modeling approaches were implemented and were applied to Chinese and Japanese text classification. To investigate the influence of word segmentation, different word segmentation approaches were investigated and applied to Chinese text. A segmentation-based approach was compared with the non-segmentation-based approach. Findings - There were two findings: the experiments show that statistical language modeling can significantly outperform standard techniques, given the same set of features; and it was found that classification with word level features normally yields improved classification performance, but that classification performance is not monotonically related to segmentation accuracy. In particular, classification performance may initially improve with increased segmentation accuracy, but eventually classification performance stops improving, and can in fact even decrease, after a certain level of segmentation accuracy. Practical implications - Apply the findings to real web text classification is ongoing work. Originality/value - The paper is very relevant to Chinese and Japanese information processing, e.g. webpage classification, web search.
Theme
Computerlinguistik
Automatisches Klassifizieren

Similar documents (author)

  1. Huang, X.; Peng, F,; An, A.; Schuurmans, D.: Dynamic Web log session identification with statistical language models (2004) 3.61
    3.6140032 = sum of:
      3.6140032 = sum of:
        1.2018057 = weight(author_txt:huang in 4094) [ClassicSimilarity], result of:
          1.2018057 = score(doc=4094,freq=1.0), product of:
            0.53210676 = queryWeight, product of:
              7.2274556 = idf(docFreq=85, maxDocs=43556)
              0.07362297 = queryNorm
            2.25858 = fieldWeight in 4094, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              7.2274556 = idf(docFreq=85, maxDocs=43556)
              0.3125 = fieldNorm(doc=4094)
        2.4121976 = weight(author_txt:peng in 4094) [ClassicSimilarity], result of:
          2.4121976 = score(doc=4094,freq=1.0), product of:
            0.84667724 = queryWeight, product of:
              1.2614195 = boost
              9.116854 = idf(docFreq=12, maxDocs=43556)
              0.07362297 = queryNorm
            2.8490167 = fieldWeight in 4094, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              9.116854 = idf(docFreq=12, maxDocs=43556)
              0.3125 = fieldNorm(doc=4094)
    
  2. Choi, B.; Peng, X.: Dynamic and hierarchical classification of Web pages (2004) 1.93
    1.9297582 = sum of:
      1.9297582 = product of:
        3.8595164 = sum of:
          3.8595164 = weight(author_txt:peng in 4553) [ClassicSimilarity], result of:
            3.8595164 = score(doc=4553,freq=1.0), product of:
              0.84667724 = queryWeight, product of:
                1.2614195 = boost
                9.116854 = idf(docFreq=12, maxDocs=43556)
                0.07362297 = queryNorm
              4.558427 = fieldWeight in 4553, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.116854 = idf(docFreq=12, maxDocs=43556)
                0.5 = fieldNorm(doc=4553)
        0.5 = coord(1/2)
    
  3. Peng, T.-Q.; Zhu, J.J.H.: Where you publish matters most : a multilevel analysis of factors affecting citations of internet studies (2012) 1.69
    1.6885384 = sum of:
      1.6885384 = product of:
        3.3770769 = sum of:
          3.3770769 = weight(author_txt:peng in 2384) [ClassicSimilarity], result of:
            3.3770769 = score(doc=2384,freq=1.0), product of:
              0.84667724 = queryWeight, product of:
                1.2614195 = boost
                9.116854 = idf(docFreq=12, maxDocs=43556)
                0.07362297 = queryNorm
              3.9886236 = fieldWeight in 2384, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.116854 = idf(docFreq=12, maxDocs=43556)
                0.4375 = fieldNorm(doc=2384)
        0.5 = coord(1/2)
    
  4. Mukhopadhyay, S.; Peng, S.; Raje, R.; Palakal, M.; Mostafa, J.: Multi-agent information classification using dynamic acquaintance lists (2003) 1.21
    1.2060988 = sum of:
      1.2060988 = product of:
        2.4121976 = sum of:
          2.4121976 = weight(author_txt:peng in 2753) [ClassicSimilarity], result of:
            2.4121976 = score(doc=2753,freq=1.0), product of:
              0.84667724 = queryWeight, product of:
                1.2614195 = boost
                9.116854 = idf(docFreq=12, maxDocs=43556)
                0.07362297 = queryNorm
              2.8490167 = fieldWeight in 2753, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.116854 = idf(docFreq=12, maxDocs=43556)
                0.3125 = fieldNorm(doc=2753)
        0.5 = coord(1/2)
    
  5. Mukhopadhyay, S.; Peng, S.; Raje, R.; Mostafa, J.; Palakal, M.: Distributed multi-agent information filtering : a comparative study (2005) 1.21
    1.2060988 = sum of:
      1.2060988 = product of:
        2.4121976 = sum of:
          2.4121976 = weight(author_txt:peng in 4557) [ClassicSimilarity], result of:
            2.4121976 = score(doc=4557,freq=1.0), product of:
              0.84667724 = queryWeight, product of:
                1.2614195 = boost
                9.116854 = idf(docFreq=12, maxDocs=43556)
                0.07362297 = queryNorm
              2.8490167 = fieldWeight in 4557, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.116854 = idf(docFreq=12, maxDocs=43556)
                0.3125 = fieldNorm(doc=4557)
        0.5 = coord(1/2)
    

Similar documents (content)

  1. Yang, C.C.; Li, K.W.: ¬A heuristic method based on a statistical approach for chinese text segmentation (2005) 0.46
    0.45620292 = sum of:
      0.45620292 = product of:
        1.6292962 = sum of:
          0.013985778 = weight(abstract_txt:based in 578) [ClassicSimilarity], result of:
            0.013985778 = score(doc=578,freq=2.0), product of:
              0.049449936 = queryWeight, product of:
                1.0710397 = boost
                3.1998224 = idf(docFreq=4826, maxDocs=43556)
                0.014428936 = queryNorm
              0.28282702 = fieldWeight in 578, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.1998224 = idf(docFreq=4826, maxDocs=43556)
                0.0625 = fieldNorm(doc=578)
          0.021336995 = weight(abstract_txt:approach in 578) [ClassicSimilarity], result of:
            0.021336995 = score(doc=578,freq=1.0), product of:
              0.09087681 = queryWeight, product of:
                1.676558 = boost
                3.7566452 = idf(docFreq=2765, maxDocs=43556)
                0.014428936 = queryNorm
              0.23479033 = fieldWeight in 578, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.7566452 = idf(docFreq=2765, maxDocs=43556)
                0.0625 = fieldNorm(doc=578)
          0.040125016 = weight(abstract_txt:performance in 578) [ClassicSimilarity], result of:
            0.040125016 = score(doc=578,freq=1.0), product of:
              0.1384547 = queryWeight, product of:
                2.069407 = boost
                4.6368976 = idf(docFreq=1146, maxDocs=43556)
                0.014428936 = queryNorm
              0.2898061 = fieldWeight in 578, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6368976 = idf(docFreq=1146, maxDocs=43556)
                0.0625 = fieldNorm(doc=578)
          0.09161017 = weight(abstract_txt:word in 578) [ClassicSimilarity], result of:
            0.09161017 = score(doc=578,freq=2.0), product of:
              0.19053876 = queryWeight, product of:
                2.4276369 = boost
                5.4395795 = idf(docFreq=513, maxDocs=43556)
                0.014428936 = queryNorm
              0.48079544 = fieldWeight in 578, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.4395795 = idf(docFreq=513, maxDocs=43556)
                0.0625 = fieldNorm(doc=578)
          0.10605852 = weight(abstract_txt:text in 578) [ClassicSimilarity], result of:
            0.10605852 = score(doc=578,freq=7.0), product of:
              0.15838924 = queryWeight, product of:
                2.710819 = boost
                4.0494018 = idf(docFreq=2063, maxDocs=43556)
                0.014428936 = queryNorm
              0.66960686 = fieldWeight in 578, product of:
                2.6457512 = tf(freq=7.0), with freq of:
                  7.0 = termFreq=7.0
                4.0494018 = idf(docFreq=2063, maxDocs=43556)
                0.0625 = fieldNorm(doc=578)
          0.3062342 = weight(abstract_txt:chinese in 578) [ClassicSimilarity], result of:
            0.3062342 = score(doc=578,freq=9.0), product of:
              0.2580195 = queryWeight, product of:
                2.824999 = boost
                6.3299446 = idf(docFreq=210, maxDocs=43556)
                0.014428936 = queryNorm
              1.1868646 = fieldWeight in 578, product of:
                3.0 = tf(freq=9.0), with freq of:
                  9.0 = termFreq=9.0
                6.3299446 = idf(docFreq=210, maxDocs=43556)
                0.0625 = fieldNorm(doc=578)
          1.0499455 = weight(abstract_txt:segmentation in 578) [ClassicSimilarity], result of:
            1.0499455 = score(doc=578,freq=9.0), product of:
              0.70698017 = queryWeight, product of:
                6.186068 = boost
                7.920603 = idf(docFreq=42, maxDocs=43556)
                0.014428936 = queryNorm
              1.485113 = fieldWeight in 578, product of:
                3.0 = tf(freq=9.0), with freq of:
                  9.0 = termFreq=9.0
                7.920603 = idf(docFreq=42, maxDocs=43556)
                0.0625 = fieldNorm(doc=578)
        0.28 = coord(7/25)
    
  2. Wang, F.L.; Yang, C.C.: Mining Web data for Chinese segmentation (2007) 0.38
    0.379801 = sum of:
      0.379801 = product of:
        1.5825043 = sum of:
          0.009889439 = weight(abstract_txt:based in 2602) [ClassicSimilarity], result of:
            0.009889439 = score(doc=2602,freq=1.0), product of:
              0.049449936 = queryWeight, product of:
                1.0710397 = boost
                3.1998224 = idf(docFreq=4826, maxDocs=43556)
                0.014428936 = queryNorm
              0.1999889 = fieldWeight in 2602, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.1998224 = idf(docFreq=4826, maxDocs=43556)
                0.0625 = fieldNorm(doc=2602)
          0.06410536 = weight(abstract_txt:language in 2602) [ClassicSimilarity], result of:
            0.06410536 = score(doc=2602,freq=3.0), product of:
              0.14132643 = queryWeight, product of:
                2.3375385 = boost
                4.1901574 = idf(docFreq=1792, maxDocs=43556)
                0.014428936 = queryNorm
              0.45359784 = fieldWeight in 2602, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.1901574 = idf(docFreq=1792, maxDocs=43556)
                0.0625 = fieldNorm(doc=2602)
          0.09161017 = weight(abstract_txt:word in 2602) [ClassicSimilarity], result of:
            0.09161017 = score(doc=2602,freq=2.0), product of:
              0.19053876 = queryWeight, product of:
                2.4276369 = boost
                5.4395795 = idf(docFreq=513, maxDocs=43556)
                0.014428936 = queryNorm
              0.48079544 = fieldWeight in 2602, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.4395795 = idf(docFreq=513, maxDocs=43556)
                0.0625 = fieldNorm(doc=2602)
          0.040086355 = weight(abstract_txt:text in 2602) [ClassicSimilarity], result of:
            0.040086355 = score(doc=2602,freq=1.0), product of:
              0.15838924 = queryWeight, product of:
                2.710819 = boost
                4.0494018 = idf(docFreq=2063, maxDocs=43556)
                0.014428936 = queryNorm
              0.2530876 = fieldWeight in 2602, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.0494018 = idf(docFreq=2063, maxDocs=43556)
                0.0625 = fieldNorm(doc=2602)
          0.2700732 = weight(abstract_txt:chinese in 2602) [ClassicSimilarity], result of:
            0.2700732 = score(doc=2602,freq=7.0), product of:
              0.2580195 = queryWeight, product of:
                2.824999 = boost
                6.3299446 = idf(docFreq=210, maxDocs=43556)
                0.014428936 = queryNorm
              1.0467162 = fieldWeight in 2602, product of:
                2.6457512 = tf(freq=7.0), with freq of:
                  7.0 = termFreq=7.0
                6.3299446 = idf(docFreq=210, maxDocs=43556)
                0.0625 = fieldNorm(doc=2602)
          1.1067398 = weight(abstract_txt:segmentation in 2602) [ClassicSimilarity], result of:
            1.1067398 = score(doc=2602,freq=10.0), product of:
              0.70698017 = queryWeight, product of:
                6.186068 = boost
                7.920603 = idf(docFreq=42, maxDocs=43556)
                0.014428936 = queryNorm
              1.5654466 = fieldWeight in 2602, product of:
                3.1622777 = tf(freq=10.0), with freq of:
                  10.0 = termFreq=10.0
                7.920603 = idf(docFreq=42, maxDocs=43556)
                0.0625 = fieldNorm(doc=2602)
        0.24 = coord(6/25)
    
  3. Huang, X.; Robertson, S.E.: Application of probilistic methods to Chinese text retrieval (1997) 0.37
    0.37406242 = sum of:
      0.37406242 = product of:
        1.1689451 = sum of:
          0.02716619 = weight(abstract_txt:purpose in 704) [ClassicSimilarity], result of:
            0.02716619 = score(doc=704,freq=1.0), product of:
              0.06466152 = queryWeight, product of:
                4.481378 = idf(docFreq=1339, maxDocs=43556)
                0.014428936 = queryNorm
              0.42012918 = fieldWeight in 704, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.481378 = idf(docFreq=1339, maxDocs=43556)
                0.09375 = fieldNorm(doc=704)
          0.020978667 = weight(abstract_txt:based in 704) [ClassicSimilarity], result of:
            0.020978667 = score(doc=704,freq=2.0), product of:
              0.049449936 = queryWeight, product of:
                1.0710397 = boost
                3.1998224 = idf(docFreq=4826, maxDocs=43556)
                0.014428936 = queryNorm
              0.42424053 = fieldWeight in 704, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.1998224 = idf(docFreq=4826, maxDocs=43556)
                0.09375 = fieldNorm(doc=704)
          0.033541318 = weight(abstract_txt:applied in 704) [ClassicSimilarity], result of:
            0.033541318 = score(doc=704,freq=1.0), product of:
              0.074418366 = queryWeight, product of:
                1.072796 = boost
                4.8076043 = idf(docFreq=966, maxDocs=43556)
                0.014428936 = queryNorm
              0.45071292 = fieldWeight in 704, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.8076043 = idf(docFreq=966, maxDocs=43556)
                0.09375 = fieldNorm(doc=704)
          0.055516876 = weight(abstract_txt:language in 704) [ClassicSimilarity], result of:
            0.055516876 = score(doc=704,freq=1.0), product of:
              0.14132643 = queryWeight, product of:
                2.3375385 = boost
                4.1901574 = idf(docFreq=1792, maxDocs=43556)
                0.014428936 = queryNorm
              0.39282727 = fieldWeight in 704, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.1901574 = idf(docFreq=1792, maxDocs=43556)
                0.09375 = fieldNorm(doc=704)
          0.13741526 = weight(abstract_txt:word in 704) [ClassicSimilarity], result of:
            0.13741526 = score(doc=704,freq=2.0), product of:
              0.19053876 = queryWeight, product of:
                2.4276369 = boost
                5.4395795 = idf(docFreq=513, maxDocs=43556)
                0.014428936 = queryNorm
              0.7211932 = fieldWeight in 704, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.4395795 = idf(docFreq=513, maxDocs=43556)
                0.09375 = fieldNorm(doc=704)
          0.104147404 = weight(abstract_txt:text in 704) [ClassicSimilarity], result of:
            0.104147404 = score(doc=704,freq=3.0), product of:
              0.15838924 = queryWeight, product of:
                2.710819 = boost
                4.0494018 = idf(docFreq=2063, maxDocs=43556)
                0.014428936 = queryNorm
              0.6575409 = fieldWeight in 704, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.0494018 = idf(docFreq=2063, maxDocs=43556)
                0.09375 = fieldNorm(doc=704)
          0.2652066 = weight(abstract_txt:chinese in 704) [ClassicSimilarity], result of:
            0.2652066 = score(doc=704,freq=3.0), product of:
              0.2580195 = queryWeight, product of:
                2.824999 = boost
                6.3299446 = idf(docFreq=210, maxDocs=43556)
                0.014428936 = queryNorm
              1.0278549 = fieldWeight in 704, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                6.3299446 = idf(docFreq=210, maxDocs=43556)
                0.09375 = fieldNorm(doc=704)
          0.52497274 = weight(abstract_txt:segmentation in 704) [ClassicSimilarity], result of:
            0.52497274 = score(doc=704,freq=1.0), product of:
              0.70698017 = queryWeight, product of:
                6.186068 = boost
                7.920603 = idf(docFreq=42, maxDocs=43556)
                0.014428936 = queryNorm
              0.7425565 = fieldWeight in 704, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.920603 = idf(docFreq=42, maxDocs=43556)
                0.09375 = fieldNorm(doc=704)
        0.32 = coord(8/25)
    
  4. Lee, K.H.; Ng, M.K.M.; Lu, Q.: Text segmentation for Chinese spell checking (1999) 0.34
    0.34332582 = sum of:
      0.34332582 = product of:
        1.2261636 = sum of:
          0.026175838 = weight(abstract_txt:level in 4911) [ClassicSimilarity], result of:
            0.026175838 = score(doc=4911,freq=2.0), product of:
              0.06560616 = queryWeight, product of:
                1.0072781 = boost
                4.5139937 = idf(docFreq=1296, maxDocs=43556)
                0.014428936 = queryNorm
              0.39898443 = fieldWeight in 4911, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.5139937 = idf(docFreq=1296, maxDocs=43556)
                0.0625 = fieldNorm(doc=4911)
          0.013985778 = weight(abstract_txt:based in 4911) [ClassicSimilarity], result of:
            0.013985778 = score(doc=4911,freq=2.0), product of:
              0.049449936 = queryWeight, product of:
                1.0710397 = boost
                3.1998224 = idf(docFreq=4826, maxDocs=43556)
                0.014428936 = queryNorm
              0.28282702 = fieldWeight in 4911, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.1998224 = idf(docFreq=4826, maxDocs=43556)
                0.0625 = fieldNorm(doc=4911)
          0.037011247 = weight(abstract_txt:language in 4911) [ClassicSimilarity], result of:
            0.037011247 = score(doc=4911,freq=1.0), product of:
              0.14132643 = queryWeight, product of:
                2.3375385 = boost
                4.1901574 = idf(docFreq=1792, maxDocs=43556)
                0.014428936 = queryNorm
              0.26188484 = fieldWeight in 4911, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.1901574 = idf(docFreq=1792, maxDocs=43556)
                0.0625 = fieldNorm(doc=4911)
          0.12955634 = weight(abstract_txt:word in 4911) [ClassicSimilarity], result of:
            0.12955634 = score(doc=4911,freq=4.0), product of:
              0.19053876 = queryWeight, product of:
                2.4276369 = boost
                5.4395795 = idf(docFreq=513, maxDocs=43556)
                0.014428936 = queryNorm
              0.67994744 = fieldWeight in 4911, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                5.4395795 = idf(docFreq=513, maxDocs=43556)
                0.0625 = fieldNorm(doc=4911)
          0.0694316 = weight(abstract_txt:text in 4911) [ClassicSimilarity], result of:
            0.0694316 = score(doc=4911,freq=3.0), product of:
              0.15838924 = queryWeight, product of:
                2.710819 = boost
                4.0494018 = idf(docFreq=2063, maxDocs=43556)
                0.014428936 = queryNorm
              0.4383606 = fieldWeight in 4911, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.0494018 = idf(docFreq=2063, maxDocs=43556)
                0.0625 = fieldNorm(doc=4911)
          0.2500392 = weight(abstract_txt:chinese in 4911) [ClassicSimilarity], result of:
            0.2500392 = score(doc=4911,freq=6.0), product of:
              0.2580195 = queryWeight, product of:
                2.824999 = boost
                6.3299446 = idf(docFreq=210, maxDocs=43556)
                0.014428936 = queryNorm
              0.9690709 = fieldWeight in 4911, product of:
                2.4494898 = tf(freq=6.0), with freq of:
                  6.0 = termFreq=6.0
                6.3299446 = idf(docFreq=210, maxDocs=43556)
                0.0625 = fieldNorm(doc=4911)
          0.6999636 = weight(abstract_txt:segmentation in 4911) [ClassicSimilarity], result of:
            0.6999636 = score(doc=4911,freq=4.0), product of:
              0.70698017 = queryWeight, product of:
                6.186068 = boost
                7.920603 = idf(docFreq=42, maxDocs=43556)
                0.014428936 = queryNorm
              0.99007535 = fieldWeight in 4911, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                7.920603 = idf(docFreq=42, maxDocs=43556)
                0.0625 = fieldNorm(doc=4911)
        0.28 = coord(7/25)
    
  5. Doval, Y.; Gómez-Rodríguez, C.: Comparing neural- and N-gram-based language models for word segmentation (2019) 0.31
    0.3088149 = sum of:
      0.3088149 = product of:
        0.8578192 = sum of:
          0.018509112 = weight(abstract_txt:level in 961) [ClassicSimilarity], result of:
            0.018509112 = score(doc=961,freq=1.0), product of:
              0.06560616 = queryWeight, product of:
                1.0072781 = boost
                4.5139937 = idf(docFreq=1296, maxDocs=43556)
                0.014428936 = queryNorm
              0.2821246 = fieldWeight in 961, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.5139937 = idf(docFreq=1296, maxDocs=43556)
                0.0625 = fieldNorm(doc=961)
          0.009889439 = weight(abstract_txt:based in 961) [ClassicSimilarity], result of:
            0.009889439 = score(doc=961,freq=1.0), product of:
              0.049449936 = queryWeight, product of:
                1.0710397 = boost
                3.1998224 = idf(docFreq=4826, maxDocs=43556)
                0.014428936 = queryNorm
              0.1999889 = fieldWeight in 961, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.1998224 = idf(docFreq=4826, maxDocs=43556)
                0.0625 = fieldNorm(doc=961)
          0.023969498 = weight(abstract_txt:task in 961) [ClassicSimilarity], result of:
            0.023969498 = score(doc=961,freq=1.0), product of:
              0.07794594 = queryWeight, product of:
                1.0979279 = boost
                4.92023 = idf(docFreq=863, maxDocs=43556)
                0.014428936 = queryNorm
              0.30751437 = fieldWeight in 961, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.92023 = idf(docFreq=863, maxDocs=43556)
                0.0625 = fieldNorm(doc=961)
          0.021336995 = weight(abstract_txt:approach in 961) [ClassicSimilarity], result of:
            0.021336995 = score(doc=961,freq=1.0), product of:
              0.09087681 = queryWeight, product of:
                1.676558 = boost
                3.7566452 = idf(docFreq=2765, maxDocs=43556)
                0.014428936 = queryNorm
              0.23479033 = fieldWeight in 961, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.7566452 = idf(docFreq=2765, maxDocs=43556)
                0.0625 = fieldNorm(doc=961)
          0.040125016 = weight(abstract_txt:performance in 961) [ClassicSimilarity], result of:
            0.040125016 = score(doc=961,freq=1.0), product of:
              0.1384547 = queryWeight, product of:
                2.069407 = boost
                4.6368976 = idf(docFreq=1146, maxDocs=43556)
                0.014428936 = queryNorm
              0.2898061 = fieldWeight in 961, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6368976 = idf(docFreq=1146, maxDocs=43556)
                0.0625 = fieldNorm(doc=961)
          0.06410536 = weight(abstract_txt:language in 961) [ClassicSimilarity], result of:
            0.06410536 = score(doc=961,freq=3.0), product of:
              0.14132643 = queryWeight, product of:
                2.3375385 = boost
                4.1901574 = idf(docFreq=1792, maxDocs=43556)
                0.014428936 = queryNorm
              0.45359784 = fieldWeight in 961, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.1901574 = idf(docFreq=1792, maxDocs=43556)
                0.0625 = fieldNorm(doc=961)
          0.1448484 = weight(abstract_txt:word in 961) [ClassicSimilarity], result of:
            0.1448484 = score(doc=961,freq=5.0), product of:
              0.19053876 = queryWeight, product of:
                2.4276369 = boost
                5.4395795 = idf(docFreq=513, maxDocs=43556)
                0.014428936 = queryNorm
              0.7602044 = fieldWeight in 961, product of:
                2.236068 = tf(freq=5.0), with freq of:
                  5.0 = termFreq=5.0
                5.4395795 = idf(docFreq=513, maxDocs=43556)
                0.0625 = fieldNorm(doc=961)
          0.040086355 = weight(abstract_txt:text in 961) [ClassicSimilarity], result of:
            0.040086355 = score(doc=961,freq=1.0), product of:
              0.15838924 = queryWeight, product of:
                2.710819 = boost
                4.0494018 = idf(docFreq=2063, maxDocs=43556)
                0.014428936 = queryNorm
              0.2530876 = fieldWeight in 961, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.0494018 = idf(docFreq=2063, maxDocs=43556)
                0.0625 = fieldNorm(doc=961)
          0.494949 = weight(abstract_txt:segmentation in 961) [ClassicSimilarity], result of:
            0.494949 = score(doc=961,freq=2.0), product of:
              0.70698017 = queryWeight, product of:
                6.186068 = boost
                7.920603 = idf(docFreq=42, maxDocs=43556)
                0.014428936 = queryNorm
              0.700089 = fieldWeight in 961, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                7.920603 = idf(docFreq=42, maxDocs=43556)
                0.0625 = fieldNorm(doc=961)
        0.36 = coord(9/25)