Document (#32832)

Author
Peng, F.
Huang, X.
Title
Machine learning for Asian language text classification
Source
Journal of documentation. 63(2007) no.3, S.378-397
Year
2007
Abstract
Purpose - The purpose of this research is to compare several machine learning techniques on the task of Asian language text classification, such as Chinese and Japanese where no word boundary information is available in written text. The paper advocates a simple language modeling based approach for this task. Design/methodology/approach - Naïve Bayes, maximum entropy model, support vector machines, and language modeling approaches were implemented and were applied to Chinese and Japanese text classification. To investigate the influence of word segmentation, different word segmentation approaches were investigated and applied to Chinese text. A segmentation-based approach was compared with the non-segmentation-based approach. Findings - There were two findings: the experiments show that statistical language modeling can significantly outperform standard techniques, given the same set of features; and it was found that classification with word level features normally yields improved classification performance, but that classification performance is not monotonically related to segmentation accuracy. In particular, classification performance may initially improve with increased segmentation accuracy, but eventually classification performance stops improving, and can in fact even decrease, after a certain level of segmentation accuracy. Practical implications - Apply the findings to real web text classification is ongoing work. Originality/value - The paper is very relevant to Chinese and Japanese information processing, e.g. webpage classification, web search.
Theme
Computerlinguistik
Automatisches Klassifizieren

Similar documents (author)

  1. Huang, X.; Peng, F,; An, A.; Schuurmans, D.: Dynamic Web log session identification with statistical language models (2004) 3.58
    3.5758846 = sum of:
      3.5758846 = sum of:
        1.2058539 = weight(author_txt:huang in 4096) [ClassicSimilarity], result of:
          1.2058539 = score(doc=4096,freq=1.0), product of:
            0.537452 = queryWeight, product of:
              7.179679 = idf(docFreq=91, maxDocs=44421)
              0.074857384 = queryNorm
            2.2436497 = fieldWeight in 4096, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              7.179679 = idf(docFreq=91, maxDocs=44421)
              0.3125 = fieldNorm(doc=4096)
        2.3700306 = weight(author_txt:peng in 4096) [ClassicSimilarity], result of:
          2.3700306 = score(doc=4096,freq=1.0), product of:
            0.8432943 = queryWeight, product of:
              1.2526212 = boost
              8.993418 = idf(docFreq=14, maxDocs=44421)
              0.074857384 = queryNorm
            2.810443 = fieldWeight in 4096, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              8.993418 = idf(docFreq=14, maxDocs=44421)
              0.3125 = fieldNorm(doc=4096)
    
  2. Choi, B.; Peng, X.: Dynamic and hierarchical classification of Web pages (2004) 1.90
    1.8960246 = sum of:
      1.8960246 = product of:
        3.7920492 = sum of:
          3.7920492 = weight(author_txt:peng in 3555) [ClassicSimilarity], result of:
            3.7920492 = score(doc=3555,freq=1.0), product of:
              0.8432943 = queryWeight, product of:
                1.2526212 = boost
                8.993418 = idf(docFreq=14, maxDocs=44421)
                0.074857384 = queryNorm
              4.496709 = fieldWeight in 3555, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.993418 = idf(docFreq=14, maxDocs=44421)
                0.5 = fieldNorm(doc=3555)
        0.5 = coord(1/2)
    
  3. Peng, X.; Shi, J.: ¬The drivers, features, and influence of first scientific collaboration among core scholars from Chinese library and information field (2024) 1.90
    1.8960246 = sum of:
      1.8960246 = product of:
        3.7920492 = sum of:
          3.7920492 = weight(author_txt:peng in 2372) [ClassicSimilarity], result of:
            3.7920492 = score(doc=2372,freq=1.0), product of:
              0.8432943 = queryWeight, product of:
                1.2526212 = boost
                8.993418 = idf(docFreq=14, maxDocs=44421)
                0.074857384 = queryNorm
              4.496709 = fieldWeight in 2372, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.993418 = idf(docFreq=14, maxDocs=44421)
                0.5 = fieldNorm(doc=2372)
        0.5 = coord(1/2)
    
  4. Peng, T.-Q.; Zhu, J.J.H.: Where you publish matters most : a multilevel analysis of factors affecting citations of internet studies (2012) 1.66
    1.6590215 = sum of:
      1.6590215 = product of:
        3.318043 = sum of:
          3.318043 = weight(author_txt:peng in 1386) [ClassicSimilarity], result of:
            3.318043 = score(doc=1386,freq=1.0), product of:
              0.8432943 = queryWeight, product of:
                1.2526212 = boost
                8.993418 = idf(docFreq=14, maxDocs=44421)
                0.074857384 = queryNorm
              3.9346204 = fieldWeight in 1386, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.993418 = idf(docFreq=14, maxDocs=44421)
                0.4375 = fieldNorm(doc=1386)
        0.5 = coord(1/2)
    
  5. Huang, G.W.: Accessing information in an information society (1989) 1.21
    1.2058539 = sum of:
      1.2058539 = product of:
        2.4117079 = sum of:
          2.4117079 = weight(author_txt:huang in 2565) [ClassicSimilarity], result of:
            2.4117079 = score(doc=2565,freq=1.0), product of:
              0.537452 = queryWeight, product of:
                7.179679 = idf(docFreq=91, maxDocs=44421)
                0.074857384 = queryNorm
              4.4872994 = fieldWeight in 2565, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.179679 = idf(docFreq=91, maxDocs=44421)
                0.625 = fieldNorm(doc=2565)
        0.5 = coord(1/2)
    

Similar documents (content)

  1. Yang, C.C.; Li, K.W.: ¬A heuristic method based on a statistical approach for chinese text segmentation (2005) 0.46
    0.45669484 = sum of:
      0.45669484 = product of:
        1.631053 = sum of:
          0.01376805 = weight(abstract_txt:based in 5580) [ClassicSimilarity], result of:
            0.01376805 = score(doc=5580,freq=2.0), product of:
              0.048936233 = queryWeight, product of:
                1.0696561 = boost
                3.1830752 = idf(docFreq=5005, maxDocs=44421)
                0.014372736 = queryNorm
              0.28134674 = fieldWeight in 5580, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.1830752 = idf(docFreq=5005, maxDocs=44421)
                0.0625 = fieldNorm(doc=5580)
          0.021075059 = weight(abstract_txt:approach in 5580) [ClassicSimilarity], result of:
            0.021075059 = score(doc=5580,freq=1.0), product of:
              0.09013311 = queryWeight, product of:
                1.6762564 = boost
                3.741144 = idf(docFreq=2864, maxDocs=44421)
                0.014372736 = queryNorm
              0.2338215 = fieldWeight in 5580, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.741144 = idf(docFreq=2864, maxDocs=44421)
                0.0625 = fieldNorm(doc=5580)
          0.039683823 = weight(abstract_txt:performance in 5580) [ClassicSimilarity], result of:
            0.039683823 = score(doc=5580,freq=1.0), product of:
              0.13744032 = queryWeight, product of:
                2.0699286 = boost
                4.619759 = idf(docFreq=1189, maxDocs=44421)
                0.014372736 = queryNorm
              0.28873494 = fieldWeight in 5580, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.619759 = idf(docFreq=1189, maxDocs=44421)
                0.0625 = fieldNorm(doc=5580)
          0.091538705 = weight(abstract_txt:word in 5580) [ClassicSimilarity], result of:
            0.091538705 = score(doc=5580,freq=2.0), product of:
              0.190443 = queryWeight, product of:
                2.4365807 = boost
                5.4380693 = idf(docFreq=524, maxDocs=44421)
                0.014372736 = queryNorm
              0.48066196 = fieldWeight in 5580, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.4380693 = idf(docFreq=524, maxDocs=44421)
                0.0625 = fieldNorm(doc=5580)
          0.10539604 = weight(abstract_txt:text in 5580) [ClassicSimilarity], result of:
            0.10539604 = score(doc=5580,freq=7.0), product of:
              0.15773174 = queryWeight, product of:
                2.7158356 = boost
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.014372736 = queryNorm
              0.66819805 = fieldWeight in 5580, product of:
                2.6457512 = tf(freq=7.0), with freq of:
                  7.0 = termFreq=7.0
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.0625 = fieldNorm(doc=5580)
          0.30175066 = weight(abstract_txt:chinese in 5580) [ClassicSimilarity], result of:
            0.30175066 = score(doc=5580,freq=9.0), product of:
              0.25549933 = queryWeight, product of:
                2.822235 = boost
                6.2987905 = idf(docFreq=221, maxDocs=44421)
                0.014372736 = queryNorm
              1.1810232 = fieldWeight in 5580, product of:
                3.0 = tf(freq=9.0), with freq of:
                  9.0 = termFreq=9.0
                6.2987905 = idf(docFreq=221, maxDocs=44421)
                0.0625 = fieldNorm(doc=5580)
          1.0578406 = weight(abstract_txt:segmentation in 5580) [ClassicSimilarity], result of:
            1.0578406 = score(doc=5580,freq=9.0), product of:
              0.71053225 = queryWeight, product of:
                6.2260013 = boost
                7.9402676 = idf(docFreq=42, maxDocs=44421)
                0.014372736 = queryNorm
              1.4888002 = fieldWeight in 5580, product of:
                3.0 = tf(freq=9.0), with freq of:
                  9.0 = termFreq=9.0
                7.9402676 = idf(docFreq=42, maxDocs=44421)
                0.0625 = fieldNorm(doc=5580)
        0.28 = coord(7/25)
    
  2. Wang, F.L.; Yang, C.C.: Mining Web data for Chinese segmentation (2007) 0.38
    0.38052574 = sum of:
      0.38052574 = product of:
        1.585524 = sum of:
          0.009735482 = weight(abstract_txt:based in 1604) [ClassicSimilarity], result of:
            0.009735482 = score(doc=1604,freq=1.0), product of:
              0.048936233 = queryWeight, product of:
                1.0696561 = boost
                3.1830752 = idf(docFreq=5005, maxDocs=44421)
                0.014372736 = queryNorm
              0.1989422 = fieldWeight in 1604, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.1830752 = idf(docFreq=5005, maxDocs=44421)
                0.0625 = fieldNorm(doc=1604)
          0.06323286 = weight(abstract_txt:language in 1604) [ClassicSimilarity], result of:
            0.06323286 = score(doc=1604,freq=3.0), product of:
              0.14004362 = queryWeight, product of:
                2.3360653 = boost
                4.1709876 = idf(docFreq=1863, maxDocs=44421)
                0.014372736 = queryNorm
              0.45152265 = fieldWeight in 1604, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.1709876 = idf(docFreq=1863, maxDocs=44421)
                0.0625 = fieldNorm(doc=1604)
          0.091538705 = weight(abstract_txt:word in 1604) [ClassicSimilarity], result of:
            0.091538705 = score(doc=1604,freq=2.0), product of:
              0.190443 = queryWeight, product of:
                2.4365807 = boost
                5.4380693 = idf(docFreq=524, maxDocs=44421)
                0.014372736 = queryNorm
              0.48066196 = fieldWeight in 1604, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.4380693 = idf(docFreq=524, maxDocs=44421)
                0.0625 = fieldNorm(doc=1604)
          0.03983596 = weight(abstract_txt:text in 1604) [ClassicSimilarity], result of:
            0.03983596 = score(doc=1604,freq=1.0), product of:
              0.15773174 = queryWeight, product of:
                2.7158356 = boost
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.014372736 = queryNorm
              0.25255513 = fieldWeight in 1604, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.0625 = fieldNorm(doc=1604)
          0.26611906 = weight(abstract_txt:chinese in 1604) [ClassicSimilarity], result of:
            0.26611906 = score(doc=1604,freq=7.0), product of:
              0.25549933 = queryWeight, product of:
                2.822235 = boost
                6.2987905 = idf(docFreq=221, maxDocs=44421)
                0.014372736 = queryNorm
              1.0415646 = fieldWeight in 1604, product of:
                2.6457512 = tf(freq=7.0), with freq of:
                  7.0 = termFreq=7.0
                6.2987905 = idf(docFreq=221, maxDocs=44421)
                0.0625 = fieldNorm(doc=1604)
          1.1150619 = weight(abstract_txt:segmentation in 1604) [ClassicSimilarity], result of:
            1.1150619 = score(doc=1604,freq=10.0), product of:
              0.71053225 = queryWeight, product of:
                6.2260013 = boost
                7.9402676 = idf(docFreq=42, maxDocs=44421)
                0.014372736 = queryNorm
              1.5693332 = fieldWeight in 1604, product of:
                3.1622777 = tf(freq=10.0), with freq of:
                  10.0 = termFreq=10.0
                7.9402676 = idf(docFreq=42, maxDocs=44421)
                0.0625 = fieldNorm(doc=1604)
        0.24 = coord(6/25)
    
  3. Huang, X.; Robertson, S.E.: Application of probilistic methods to Chinese text retrieval (1997) 0.37
    0.37332004 = sum of:
      0.37332004 = product of:
        1.1666251 = sum of:
          0.026847178 = weight(abstract_txt:purpose in 5706) [ClassicSimilarity], result of:
            0.026847178 = score(doc=5706,freq=1.0), product of:
              0.06415543 = queryWeight, product of:
                4.4636893 = idf(docFreq=1390, maxDocs=44421)
                0.014372736 = queryNorm
              0.41847086 = fieldWeight in 5706, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.4636893 = idf(docFreq=1390, maxDocs=44421)
                0.09375 = fieldNorm(doc=5706)
          0.020652074 = weight(abstract_txt:based in 5706) [ClassicSimilarity], result of:
            0.020652074 = score(doc=5706,freq=2.0), product of:
              0.048936233 = queryWeight, product of:
                1.0696561 = boost
                3.1830752 = idf(docFreq=5005, maxDocs=44421)
                0.014372736 = queryNorm
              0.4220201 = fieldWeight in 5706, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.1830752 = idf(docFreq=5005, maxDocs=44421)
                0.09375 = fieldNorm(doc=5706)
          0.033315703 = weight(abstract_txt:applied in 5706) [ClassicSimilarity], result of:
            0.033315703 = score(doc=5706,freq=1.0), product of:
              0.07408557 = queryWeight, product of:
                1.0746081 = boost
                4.7967167 = idf(docFreq=996, maxDocs=44421)
                0.014372736 = queryNorm
              0.4496922 = fieldWeight in 5706, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.7967167 = idf(docFreq=996, maxDocs=44421)
                0.09375 = fieldNorm(doc=5706)
          0.054761264 = weight(abstract_txt:language in 5706) [ClassicSimilarity], result of:
            0.054761264 = score(doc=5706,freq=1.0), product of:
              0.14004362 = queryWeight, product of:
                2.3360653 = boost
                4.1709876 = idf(docFreq=1863, maxDocs=44421)
                0.014372736 = queryNorm
              0.39103007 = fieldWeight in 5706, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.1709876 = idf(docFreq=1863, maxDocs=44421)
                0.09375 = fieldNorm(doc=5706)
          0.13730805 = weight(abstract_txt:word in 5706) [ClassicSimilarity], result of:
            0.13730805 = score(doc=5706,freq=2.0), product of:
              0.190443 = queryWeight, product of:
                2.4365807 = boost
                5.4380693 = idf(docFreq=524, maxDocs=44421)
                0.014372736 = queryNorm
              0.7209929 = fieldWeight in 5706, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.4380693 = idf(docFreq=524, maxDocs=44421)
                0.09375 = fieldNorm(doc=5706)
          0.103496864 = weight(abstract_txt:text in 5706) [ClassicSimilarity], result of:
            0.103496864 = score(doc=5706,freq=3.0), product of:
              0.15773174 = queryWeight, product of:
                2.7158356 = boost
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.014372736 = queryNorm
              0.6561575 = fieldWeight in 5706, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.09375 = fieldNorm(doc=5706)
          0.26132375 = weight(abstract_txt:chinese in 5706) [ClassicSimilarity], result of:
            0.26132375 = score(doc=5706,freq=3.0), product of:
              0.25549933 = queryWeight, product of:
                2.822235 = boost
                6.2987905 = idf(docFreq=221, maxDocs=44421)
                0.014372736 = queryNorm
              1.0227962 = fieldWeight in 5706, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                6.2987905 = idf(docFreq=221, maxDocs=44421)
                0.09375 = fieldNorm(doc=5706)
          0.5289203 = weight(abstract_txt:segmentation in 5706) [ClassicSimilarity], result of:
            0.5289203 = score(doc=5706,freq=1.0), product of:
              0.71053225 = queryWeight, product of:
                6.2260013 = boost
                7.9402676 = idf(docFreq=42, maxDocs=44421)
                0.014372736 = queryNorm
              0.7444001 = fieldWeight in 5706, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.9402676 = idf(docFreq=42, maxDocs=44421)
                0.09375 = fieldNorm(doc=5706)
        0.32 = coord(8/25)
    
  4. Lee, K.H.; Ng, M.K.M.; Lu, Q.: Text segmentation for Chinese spell checking (1999) 0.34
    0.3433315 = sum of:
      0.3433315 = product of:
        1.2261839 = sum of:
          0.025849717 = weight(abstract_txt:level in 4913) [ClassicSimilarity], result of:
            0.025849717 = score(doc=4913,freq=2.0), product of:
              0.06506124 = queryWeight, product of:
                1.0070348 = boost
                4.4950905 = idf(docFreq=1347, maxDocs=44421)
                0.014372736 = queryNorm
              0.39731362 = fieldWeight in 4913, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.4950905 = idf(docFreq=1347, maxDocs=44421)
                0.0625 = fieldNorm(doc=4913)
          0.01376805 = weight(abstract_txt:based in 4913) [ClassicSimilarity], result of:
            0.01376805 = score(doc=4913,freq=2.0), product of:
              0.048936233 = queryWeight, product of:
                1.0696561 = boost
                3.1830752 = idf(docFreq=5005, maxDocs=44421)
                0.014372736 = queryNorm
              0.28134674 = fieldWeight in 4913, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.1830752 = idf(docFreq=5005, maxDocs=44421)
                0.0625 = fieldNorm(doc=4913)
          0.036507513 = weight(abstract_txt:language in 4913) [ClassicSimilarity], result of:
            0.036507513 = score(doc=4913,freq=1.0), product of:
              0.14004362 = queryWeight, product of:
                2.3360653 = boost
                4.1709876 = idf(docFreq=1863, maxDocs=44421)
                0.014372736 = queryNorm
              0.26068673 = fieldWeight in 4913, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.1709876 = idf(docFreq=1863, maxDocs=44421)
                0.0625 = fieldNorm(doc=4913)
          0.12945528 = weight(abstract_txt:word in 4913) [ClassicSimilarity], result of:
            0.12945528 = score(doc=4913,freq=4.0), product of:
              0.190443 = queryWeight, product of:
                2.4365807 = boost
                5.4380693 = idf(docFreq=524, maxDocs=44421)
                0.014372736 = queryNorm
              0.67975867 = fieldWeight in 4913, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                5.4380693 = idf(docFreq=524, maxDocs=44421)
                0.0625 = fieldNorm(doc=4913)
          0.068997905 = weight(abstract_txt:text in 4913) [ClassicSimilarity], result of:
            0.068997905 = score(doc=4913,freq=3.0), product of:
              0.15773174 = queryWeight, product of:
                2.7158356 = boost
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.014372736 = queryNorm
              0.4374383 = fieldWeight in 4913, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.0625 = fieldNorm(doc=4913)
          0.24637838 = weight(abstract_txt:chinese in 4913) [ClassicSimilarity], result of:
            0.24637838 = score(doc=4913,freq=6.0), product of:
              0.25549933 = queryWeight, product of:
                2.822235 = boost
                6.2987905 = idf(docFreq=221, maxDocs=44421)
                0.014372736 = queryNorm
              0.96430147 = fieldWeight in 4913, product of:
                2.4494898 = tf(freq=6.0), with freq of:
                  6.0 = termFreq=6.0
                6.2987905 = idf(docFreq=221, maxDocs=44421)
                0.0625 = fieldNorm(doc=4913)
          0.705227 = weight(abstract_txt:segmentation in 4913) [ClassicSimilarity], result of:
            0.705227 = score(doc=4913,freq=4.0), product of:
              0.71053225 = queryWeight, product of:
                6.2260013 = boost
                7.9402676 = idf(docFreq=42, maxDocs=44421)
                0.014372736 = queryNorm
              0.99253345 = fieldWeight in 4913, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                7.9402676 = idf(docFreq=42, maxDocs=44421)
                0.0625 = fieldNorm(doc=4913)
        0.28 = coord(7/25)
    
  5. Doval, Y.; Gómez-Rodríguez, C.: Comparing neural- and N-gram-based language models for word segmentation (2019) 0.31
    0.3092428 = sum of:
      0.3092428 = product of:
        0.8590078 = sum of:
          0.018278511 = weight(abstract_txt:level in 675) [ClassicSimilarity], result of:
            0.018278511 = score(doc=675,freq=1.0), product of:
              0.06506124 = queryWeight, product of:
                1.0070348 = boost
                4.4950905 = idf(docFreq=1347, maxDocs=44421)
                0.014372736 = queryNorm
              0.28094316 = fieldWeight in 675, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.4950905 = idf(docFreq=1347, maxDocs=44421)
                0.0625 = fieldNorm(doc=675)
          0.009735482 = weight(abstract_txt:based in 675) [ClassicSimilarity], result of:
            0.009735482 = score(doc=675,freq=1.0), product of:
              0.048936233 = queryWeight, product of:
                1.0696561 = boost
                3.1830752 = idf(docFreq=5005, maxDocs=44421)
                0.014372736 = queryNorm
              0.1989422 = fieldWeight in 675, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.1830752 = idf(docFreq=5005, maxDocs=44421)
                0.0625 = fieldNorm(doc=675)
          0.023759915 = weight(abstract_txt:task in 675) [ClassicSimilarity], result of:
            0.023759915 = score(doc=675,freq=1.0), product of:
              0.077492274 = queryWeight, product of:
                1.0990374 = boost
                4.9057617 = idf(docFreq=893, maxDocs=44421)
                0.014372736 = queryNorm
              0.3066101 = fieldWeight in 675, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.9057617 = idf(docFreq=893, maxDocs=44421)
                0.0625 = fieldNorm(doc=675)
          0.021075059 = weight(abstract_txt:approach in 675) [ClassicSimilarity], result of:
            0.021075059 = score(doc=675,freq=1.0), product of:
              0.09013311 = queryWeight, product of:
                1.6762564 = boost
                3.741144 = idf(docFreq=2864, maxDocs=44421)
                0.014372736 = queryNorm
              0.2338215 = fieldWeight in 675, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.741144 = idf(docFreq=2864, maxDocs=44421)
                0.0625 = fieldNorm(doc=675)
          0.039683823 = weight(abstract_txt:performance in 675) [ClassicSimilarity], result of:
            0.039683823 = score(doc=675,freq=1.0), product of:
              0.13744032 = queryWeight, product of:
                2.0699286 = boost
                4.619759 = idf(docFreq=1189, maxDocs=44421)
                0.014372736 = queryNorm
              0.28873494 = fieldWeight in 675, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.619759 = idf(docFreq=1189, maxDocs=44421)
                0.0625 = fieldNorm(doc=675)
          0.06323286 = weight(abstract_txt:language in 675) [ClassicSimilarity], result of:
            0.06323286 = score(doc=675,freq=3.0), product of:
              0.14004362 = queryWeight, product of:
                2.3360653 = boost
                4.1709876 = idf(docFreq=1863, maxDocs=44421)
                0.014372736 = queryNorm
              0.45152265 = fieldWeight in 675, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.1709876 = idf(docFreq=1863, maxDocs=44421)
                0.0625 = fieldNorm(doc=675)
          0.1447354 = weight(abstract_txt:word in 675) [ClassicSimilarity], result of:
            0.1447354 = score(doc=675,freq=5.0), product of:
              0.190443 = queryWeight, product of:
                2.4365807 = boost
                5.4380693 = idf(docFreq=524, maxDocs=44421)
                0.014372736 = queryNorm
              0.7599933 = fieldWeight in 675, product of:
                2.236068 = tf(freq=5.0), with freq of:
                  5.0 = termFreq=5.0
                5.4380693 = idf(docFreq=524, maxDocs=44421)
                0.0625 = fieldNorm(doc=675)
          0.03983596 = weight(abstract_txt:text in 675) [ClassicSimilarity], result of:
            0.03983596 = score(doc=675,freq=1.0), product of:
              0.15773174 = queryWeight, product of:
                2.7158356 = boost
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.014372736 = queryNorm
              0.25255513 = fieldWeight in 675, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.0625 = fieldNorm(doc=675)
          0.4986708 = weight(abstract_txt:segmentation in 675) [ClassicSimilarity], result of:
            0.4986708 = score(doc=675,freq=2.0), product of:
              0.71053225 = queryWeight, product of:
                6.2260013 = boost
                7.9402676 = idf(docFreq=42, maxDocs=44421)
                0.014372736 = queryNorm
              0.7018271 = fieldWeight in 675, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                7.9402676 = idf(docFreq=42, maxDocs=44421)
                0.0625 = fieldNorm(doc=675)
        0.36 = coord(9/25)