Document (#38443)

Author
Mesquita, L.A.P.
Souza, R.R.
Baracho Porto, R.M.A.
Title
Noun phrases in automatic indexing: : a structural analysis of the distribution of relevant terms in doctoral theses
Source
Knowledge organization in the 21st century: between historical patterns and future prospects. Proceedings of the Thirteenth International ISKO Conference 19-22 May 2014, Kraków, Poland. Ed.: Wieslaw Babik
Imprint
Würzburg : Ergon Verlag
Year
2014
Pages
S.327-334
Series
Advances in knowledge organization; vol. 14
Abstract
The main objective of this research was to analyze whether there was a characteristic distribution behavior of relevant terms over a scientific text that could contribute as a criterion for their process of automatic indexing. The terms considered in this study were only full noun phrases contained in the texts themselves. The texts were considered a total of 98 doctoral theses of the eight areas of knowledge in a same university. Initially, 20 full noun phrases were automatically extracted from each text as candidates to be the most relevant terms, and each author of each text assigned a relevance value 0-6 (not relevant and highly relevant, respectively) for each of the 20 noun phrases sent. Only, 22.1 % of noun phrases were considered not relevant. A relevance values of the terms assigned by the authors were associated with their positions in the text. Each full noun phrases found in the text was considered as a valid linear position. The results that were obtained showed values resulting from this distribution by considering two types of position: linear, with values consolidated into ten equal consecutive parts; and structural, considering parts of the text (such as introduction, development and conclusion). As a result of considerable importance, all areas of knowledge related to the Natural Sciences showed a characteristic behavior in the distribution of relevant terms, as well as all areas of knowledge related to Social Sciences showed the same characteristic behavior of distribution, but distinct from the Natural Sciences. The difference of the distribution behavior between the Natural and Social Sciences can be clearly visualized through graphs. All behaviors, including the general behavior of all areas of knowledge together, were characterized in polynomial equations and can be applied in future as criteria for automatic indexing. Until the present date this work has become inedited of for two reasons: to present a method for characterizing the distribution of relevant terms in a scientific text, and also, through this method, pointing out a quantitative trait difference between the Natural and Social Sciences.
Content
Vgl.: http://www.ergon-verlag.de/isko_ko/downloads/aiko_vol_14_2014_45.pdf.
Theme
Automatisches Indexieren

Similar documents (author)

  1. Almeida, M.B.; Souza, R.R.; Porto, R.B.: Looking for the identity of information science in the age of big data, computing clouds and social networks (2015) 4.75
    4.7480817 = sum of:
      4.7480817 = sum of:
        1.747206 = weight(author_txt:souza in 4453) [ClassicSimilarity], result of:
          1.747206 = score(doc=4453,freq=1.0), product of:
            0.57195526 = queryWeight, product of:
              8.146119 = idf(docFreq=34, maxDocs=44421)
              0.07021199 = queryNorm
            3.0547948 = fieldWeight in 4453, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              8.146119 = idf(docFreq=34, maxDocs=44421)
              0.375 = fieldNorm(doc=4453)
        3.0008757 = weight(author_txt:porto in 4453) [ClassicSimilarity], result of:
          3.0008757 = score(doc=4453,freq=1.0), product of:
            0.82028484 = queryWeight, product of:
              1.1975712 = boost
              9.755557 = idf(docFreq=6, maxDocs=44421)
              0.07021199 = queryNorm
            3.6583338 = fieldWeight in 4453, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              9.755557 = idf(docFreq=6, maxDocs=44421)
              0.375 = fieldNorm(doc=4453)
    
  2. Porto, R. => Porto, R.B.: 2.83
    2.829253 = sum of:
      2.829253 = product of:
        5.658506 = sum of:
          5.658506 = weight(author_txt:porto in 4386) [ClassicSimilarity], result of:
            5.658506 = score(doc=4386,freq=2.0), product of:
              0.82028484 = queryWeight, product of:
                1.1975712 = boost
                9.755557 = idf(docFreq=6, maxDocs=44421)
                0.07021199 = queryNorm
              6.8982205 = fieldWeight in 4386, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                9.755557 = idf(docFreq=6, maxDocs=44421)
                0.5 = fieldNorm(doc=4386)
        0.5 = coord(1/2)
    
  3. Gomez, I. Porto- => Porto-Gomez, I.: 2.12
    2.1219397 = sum of:
      2.1219397 = product of:
        4.2438793 = sum of:
          4.2438793 = weight(author_txt:porto in 372) [ClassicSimilarity], result of:
            4.2438793 = score(doc=372,freq=2.0), product of:
              0.82028484 = queryWeight, product of:
                1.1975712 = boost
                9.755557 = idf(docFreq=6, maxDocs=44421)
                0.07021199 = queryNorm
              5.1736655 = fieldWeight in 372, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                9.755557 = idf(docFreq=6, maxDocs=44421)
                0.375 = fieldNorm(doc=372)
        0.5 = coord(1/2)
    
  4. Dal Porto, S.; Marchitelli, A.: ¬The functionality and flexibility of traditional classification schemes applied to a Content Management System (CMS) : facets, DDC, JITA (2006) 1.75
    1.7505109 = sum of:
      1.7505109 = product of:
        3.5010219 = sum of:
          3.5010219 = weight(author_txt:porto in 299) [ClassicSimilarity], result of:
            3.5010219 = score(doc=299,freq=1.0), product of:
              0.82028484 = queryWeight, product of:
                1.1975712 = boost
                9.755557 = idf(docFreq=6, maxDocs=44421)
                0.07021199 = queryNorm
              4.2680564 = fieldWeight in 299, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.755557 = idf(docFreq=6, maxDocs=44421)
                0.4375 = fieldNorm(doc=299)
        0.5 = coord(1/2)
    
  5. Souza, S.d.: Informacion : utopia y realidad de la bibliotelogia (1996) 1.46
    1.4560049 = sum of:
      1.4560049 = product of:
        2.9120097 = sum of:
          2.9120097 = weight(author_txt:souza in 824) [ClassicSimilarity], result of:
            2.9120097 = score(doc=824,freq=1.0), product of:
              0.57195526 = queryWeight, product of:
                8.146119 = idf(docFreq=34, maxDocs=44421)
                0.07021199 = queryNorm
              5.0913243 = fieldWeight in 824, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.146119 = idf(docFreq=34, maxDocs=44421)
                0.625 = fieldNorm(doc=824)
        0.5 = coord(1/2)
    

Similar documents (content)

  1. Souza, R.R.; Raghavan, K.S.: ¬A methodology for noun phrase-based automatic indexing (2006) 0.32
    0.3206301 = sum of:
      0.3206301 = product of:
        1.1451075 = sum of:
          0.027214456 = weight(abstract_txt:indexing in 298) [ClassicSimilarity], result of:
            0.027214456 = score(doc=298,freq=1.0), product of:
              0.080073625 = queryWeight, product of:
                1.0613085 = boost
                4.3503094 = idf(docFreq=1557, maxDocs=44421)
                0.01734314 = queryNorm
              0.33986792 = fieldWeight in 298, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.3503094 = idf(docFreq=1557, maxDocs=44421)
                0.078125 = fieldNorm(doc=298)
          0.019657962 = weight(abstract_txt:knowledge in 298) [ClassicSimilarity], result of:
            0.019657962 = score(doc=298,freq=1.0), product of:
              0.07095151 = queryWeight, product of:
                1.1535783 = boost
                3.5463927 = idf(docFreq=3480, maxDocs=44421)
                0.01734314 = queryNorm
              0.27706194 = fieldWeight in 298, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.5463927 = idf(docFreq=3480, maxDocs=44421)
                0.078125 = fieldNorm(doc=298)
          0.05089142 = weight(abstract_txt:text in 298) [ClassicSimilarity], result of:
            0.05089142 = score(doc=298,freq=1.0), product of:
              0.16120495 = queryWeight, product of:
                2.300247 = boost
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.01734314 = queryNorm
              0.3156939 = fieldWeight in 298, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.078125 = fieldNorm(doc=298)
          0.05099842 = weight(abstract_txt:terms in 298) [ClassicSimilarity], result of:
            0.05099842 = score(doc=298,freq=1.0), product of:
              0.16143082 = queryWeight, product of:
                2.301858 = boost
                4.043712 = idf(docFreq=2116, maxDocs=44421)
                0.01734314 = queryNorm
              0.31591502 = fieldWeight in 298, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.043712 = idf(docFreq=2116, maxDocs=44421)
                0.078125 = fieldNorm(doc=298)
          0.087482505 = weight(abstract_txt:relevant in 298) [ClassicSimilarity], result of:
            0.087482505 = score(doc=298,freq=1.0), product of:
              0.2418578 = queryWeight, product of:
                3.012044 = boost
                4.6298943 = idf(docFreq=1177, maxDocs=44421)
                0.01734314 = queryNorm
              0.3617105 = fieldWeight in 298, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6298943 = idf(docFreq=1177, maxDocs=44421)
                0.078125 = fieldNorm(doc=298)
          0.37179032 = weight(abstract_txt:phrases in 298) [ClassicSimilarity], result of:
            0.37179032 = score(doc=298,freq=3.0), product of:
              0.39975265 = queryWeight, product of:
                3.3535714 = boost
                6.8731537 = idf(docFreq=124, maxDocs=44421)
                0.01734314 = queryNorm
              0.9300509 = fieldWeight in 298, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                6.8731537 = idf(docFreq=124, maxDocs=44421)
                0.078125 = fieldNorm(doc=298)
          0.5370725 = weight(abstract_txt:noun in 298) [ClassicSimilarity], result of:
            0.5370725 = score(doc=298,freq=3.0), product of:
              0.5108357 = queryWeight, product of:
                3.790989 = boost
                7.769642 = idf(docFreq=50, maxDocs=44421)
                0.01734314 = queryNorm
              1.0513605 = fieldWeight in 298, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                7.769642 = idf(docFreq=50, maxDocs=44421)
                0.078125 = fieldNorm(doc=298)
        0.28 = coord(7/25)
    
  2. Kim, W.; Wilbur, W.J.: Corpus-based statistical screening for content-bearing terms (2001) 0.26
    0.2576743 = sum of:
      0.2576743 = product of:
        0.71576196 = sum of:
          0.0549257 = weight(abstract_txt:values in 188) [ClassicSimilarity], result of:
            0.0549257 = score(doc=188,freq=2.0), product of:
              0.14267984 = queryWeight, product of:
                1.4167011 = boost
                5.807065 = idf(docFreq=362, maxDocs=44421)
                0.01734314 = queryNorm
              0.38495767 = fieldWeight in 188, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.807065 = idf(docFreq=362, maxDocs=44421)
                0.046875 = fieldNorm(doc=188)
          0.033415563 = weight(abstract_txt:considered in 188) [ClassicSimilarity], result of:
            0.033415563 = score(doc=188,freq=1.0), product of:
              0.14205864 = queryWeight, product of:
                1.6323005 = boost
                5.0181065 = idf(docFreq=798, maxDocs=44421)
                0.01734314 = queryNorm
              0.23522374 = fieldWeight in 188, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.0181065 = idf(docFreq=798, maxDocs=44421)
                0.046875 = fieldNorm(doc=188)
          0.034452107 = weight(abstract_txt:natural in 188) [ClassicSimilarity], result of:
            0.034452107 = score(doc=188,freq=1.0), product of:
              0.14498141 = queryWeight, product of:
                1.6490067 = boost
                5.0694656 = idf(docFreq=758, maxDocs=44421)
                0.01734314 = queryNorm
              0.2376312 = fieldWeight in 188, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.0694656 = idf(docFreq=758, maxDocs=44421)
                0.046875 = fieldNorm(doc=188)
          0.056530476 = weight(abstract_txt:each in 188) [ClassicSimilarity], result of:
            0.056530476 = score(doc=188,freq=6.0), product of:
              0.11956658 = queryWeight, product of:
                1.6742727 = boost
                4.1177115 = idf(docFreq=1965, maxDocs=44421)
                0.01734314 = queryNorm
              0.47279495 = fieldWeight in 188, product of:
                2.4494898 = tf(freq=6.0), with freq of:
                  6.0 = termFreq=6.0
                4.1177115 = idf(docFreq=1965, maxDocs=44421)
                0.046875 = fieldNorm(doc=188)
          0.022828234 = weight(abstract_txt:were in 188) [ClassicSimilarity], result of:
            0.022828234 = score(doc=188,freq=1.0), product of:
              0.13278918 = queryWeight, product of:
                2.087693 = boost
                3.6674848 = idf(docFreq=3083, maxDocs=44421)
                0.01734314 = queryNorm
              0.17191336 = fieldWeight in 188, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.6674848 = idf(docFreq=3083, maxDocs=44421)
                0.046875 = fieldNorm(doc=188)
          0.030534852 = weight(abstract_txt:text in 188) [ClassicSimilarity], result of:
            0.030534852 = score(doc=188,freq=1.0), product of:
              0.16120495 = queryWeight, product of:
                2.300247 = boost
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.01734314 = queryNorm
              0.18941635 = fieldWeight in 188, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.046875 = fieldNorm(doc=188)
          0.0611981 = weight(abstract_txt:terms in 188) [ClassicSimilarity], result of:
            0.0611981 = score(doc=188,freq=4.0), product of:
              0.16143082 = queryWeight, product of:
                2.301858 = boost
                4.043712 = idf(docFreq=2116, maxDocs=44421)
                0.01734314 = queryNorm
              0.379098 = fieldWeight in 188, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                4.043712 = idf(docFreq=2116, maxDocs=44421)
                0.046875 = fieldNorm(doc=188)
          0.0811255 = weight(abstract_txt:distribution in 188) [ClassicSimilarity], result of:
            0.0811255 = score(doc=188,freq=1.0), product of:
              0.30923316 = queryWeight, product of:
                3.185872 = boost
                5.5966744 = idf(docFreq=447, maxDocs=44421)
                0.01734314 = queryNorm
              0.26234412 = fieldWeight in 188, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.5966744 = idf(docFreq=447, maxDocs=44421)
                0.046875 = fieldNorm(doc=188)
          0.3407514 = weight(abstract_txt:phrases in 188) [ClassicSimilarity], result of:
            0.3407514 = score(doc=188,freq=7.0), product of:
              0.39975265 = queryWeight, product of:
                3.3535714 = boost
                6.8731537 = idf(docFreq=124, maxDocs=44421)
                0.01734314 = queryNorm
              0.85240567 = fieldWeight in 188, product of:
                2.6457512 = tf(freq=7.0), with freq of:
                  7.0 = termFreq=7.0
                6.8731537 = idf(docFreq=124, maxDocs=44421)
                0.046875 = fieldNorm(doc=188)
        0.36 = coord(9/25)
    
  3. Vlachidis, A.; Tudhope, D.: ¬A knowledge-based approach to information extraction for semantic interoperability in the archaeology domain (2016) 0.21
    0.21120866 = sum of:
      0.21120866 = product of:
        0.6600271 = sum of:
          0.030789642 = weight(abstract_txt:indexing in 3895) [ClassicSimilarity], result of:
            0.030789642 = score(doc=3895,freq=2.0), product of:
              0.080073625 = queryWeight, product of:
                1.0613085 = boost
                4.3503094 = idf(docFreq=1557, maxDocs=44421)
                0.01734314 = queryNorm
              0.38451666 = fieldWeight in 3895, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.3503094 = idf(docFreq=1557, maxDocs=44421)
                0.0625 = fieldNorm(doc=3895)
          0.015726369 = weight(abstract_txt:knowledge in 3895) [ClassicSimilarity], result of:
            0.015726369 = score(doc=3895,freq=1.0), product of:
              0.07095151 = queryWeight, product of:
                1.1535783 = boost
                3.5463927 = idf(docFreq=3480, maxDocs=44421)
                0.01734314 = queryNorm
              0.22164954 = fieldWeight in 3895, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.5463927 = idf(docFreq=3480, maxDocs=44421)
                0.0625 = fieldNorm(doc=3895)
          0.03709004 = weight(abstract_txt:automatic in 3895) [ClassicSimilarity], result of:
            0.03709004 = score(doc=3895,freq=1.0), product of:
              0.11421801 = queryWeight, product of:
                1.2675474 = boost
                5.1956835 = idf(docFreq=668, maxDocs=44421)
                0.01734314 = queryNorm
              0.32473022 = fieldWeight in 3895, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.1956835 = idf(docFreq=668, maxDocs=44421)
                0.0625 = fieldNorm(doc=3895)
          0.045936145 = weight(abstract_txt:natural in 3895) [ClassicSimilarity], result of:
            0.045936145 = score(doc=3895,freq=1.0), product of:
              0.14498141 = queryWeight, product of:
                1.6490067 = boost
                5.0694656 = idf(docFreq=758, maxDocs=44421)
                0.01734314 = queryNorm
              0.3168416 = fieldWeight in 3895, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.0694656 = idf(docFreq=758, maxDocs=44421)
                0.0625 = fieldNorm(doc=3895)
          0.04071314 = weight(abstract_txt:text in 3895) [ClassicSimilarity], result of:
            0.04071314 = score(doc=3895,freq=1.0), product of:
              0.16120495 = queryWeight, product of:
                2.300247 = boost
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.01734314 = queryNorm
              0.25255513 = fieldWeight in 3895, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.0625 = fieldNorm(doc=3895)
          0.069986 = weight(abstract_txt:relevant in 3895) [ClassicSimilarity], result of:
            0.069986 = score(doc=3895,freq=1.0), product of:
              0.2418578 = queryWeight, product of:
                3.012044 = boost
                4.6298943 = idf(docFreq=1177, maxDocs=44421)
                0.01734314 = queryNorm
              0.2893684 = fieldWeight in 3895, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6298943 = idf(docFreq=1177, maxDocs=44421)
                0.0625 = fieldNorm(doc=3895)
          0.17172259 = weight(abstract_txt:phrases in 3895) [ClassicSimilarity], result of:
            0.17172259 = score(doc=3895,freq=1.0), product of:
              0.39975265 = queryWeight, product of:
                3.3535714 = boost
                6.8731537 = idf(docFreq=124, maxDocs=44421)
                0.01734314 = queryNorm
              0.4295721 = fieldWeight in 3895, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.8731537 = idf(docFreq=124, maxDocs=44421)
                0.0625 = fieldNorm(doc=3895)
          0.24806316 = weight(abstract_txt:noun in 3895) [ClassicSimilarity], result of:
            0.24806316 = score(doc=3895,freq=1.0), product of:
              0.5108357 = queryWeight, product of:
                3.790989 = boost
                7.769642 = idf(docFreq=50, maxDocs=44421)
                0.01734314 = queryNorm
              0.48560262 = fieldWeight in 3895, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.769642 = idf(docFreq=50, maxDocs=44421)
                0.0625 = fieldNorm(doc=3895)
        0.32 = coord(8/25)
    
  4. Salles, T.; Rocha, L.; Gonçalves, M.A.; Almeida, J.M.; Mourão, F.; Meira Jr., W.; Viegas, F.: ¬A quantitative analysis of the temporal effects on automatic text classification (2016) 0.19
    0.18837036 = sum of:
      0.18837036 = product of:
        0.523251 = sum of:
          0.03158912 = weight(abstract_txt:full in 4014) [ClassicSimilarity], result of:
            0.03158912 = score(doc=4014,freq=1.0), product of:
              0.10262537 = queryWeight, product of:
                1.2015014 = boost
                4.9249606 = idf(docFreq=876, maxDocs=44421)
                0.01734314 = queryNorm
              0.30781004 = fieldWeight in 4014, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.9249606 = idf(docFreq=876, maxDocs=44421)
                0.0625 = fieldNorm(doc=4014)
          0.03709004 = weight(abstract_txt:automatic in 4014) [ClassicSimilarity], result of:
            0.03709004 = score(doc=4014,freq=1.0), product of:
              0.11421801 = queryWeight, product of:
                1.2675474 = boost
                5.1956835 = idf(docFreq=668, maxDocs=44421)
                0.01734314 = queryNorm
              0.32473022 = fieldWeight in 4014, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.1956835 = idf(docFreq=668, maxDocs=44421)
                0.0625 = fieldNorm(doc=4014)
          0.04455409 = weight(abstract_txt:considered in 4014) [ClassicSimilarity], result of:
            0.04455409 = score(doc=4014,freq=1.0), product of:
              0.14205864 = queryWeight, product of:
                1.6323005 = boost
                5.0181065 = idf(docFreq=798, maxDocs=44421)
                0.01734314 = queryNorm
              0.31363165 = fieldWeight in 4014, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.0181065 = idf(docFreq=798, maxDocs=44421)
                0.0625 = fieldNorm(doc=4014)
          0.04351718 = weight(abstract_txt:each in 4014) [ClassicSimilarity], result of:
            0.04351718 = score(doc=4014,freq=2.0), product of:
              0.11956658 = queryWeight, product of:
                1.6742727 = boost
                4.1177115 = idf(docFreq=1965, maxDocs=44421)
                0.01734314 = queryNorm
              0.3639577 = fieldWeight in 4014, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.1177115 = idf(docFreq=1965, maxDocs=44421)
                0.0625 = fieldNorm(doc=4014)
          0.062031038 = weight(abstract_txt:behavior in 4014) [ClassicSimilarity], result of:
            0.062031038 = score(doc=4014,freq=1.0), product of:
              0.19080307 = queryWeight, product of:
                2.1150174 = boost
                5.2016807 = idf(docFreq=664, maxDocs=44421)
                0.01734314 = queryNorm
              0.32510504 = fieldWeight in 4014, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.2016807 = idf(docFreq=664, maxDocs=44421)
                0.0625 = fieldNorm(doc=4014)
          0.04071314 = weight(abstract_txt:text in 4014) [ClassicSimilarity], result of:
            0.04071314 = score(doc=4014,freq=1.0), product of:
              0.16120495 = queryWeight, product of:
                2.300247 = boost
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.01734314 = queryNorm
              0.25255513 = fieldWeight in 4014, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.0625 = fieldNorm(doc=4014)
          0.040798735 = weight(abstract_txt:terms in 4014) [ClassicSimilarity], result of:
            0.040798735 = score(doc=4014,freq=1.0), product of:
              0.16143082 = queryWeight, product of:
                2.301858 = boost
                4.043712 = idf(docFreq=2116, maxDocs=44421)
                0.01734314 = queryNorm
              0.252732 = fieldWeight in 4014, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.043712 = idf(docFreq=2116, maxDocs=44421)
                0.0625 = fieldNorm(doc=4014)
          0.069986 = weight(abstract_txt:relevant in 4014) [ClassicSimilarity], result of:
            0.069986 = score(doc=4014,freq=1.0), product of:
              0.2418578 = queryWeight, product of:
                3.012044 = boost
                4.6298943 = idf(docFreq=1177, maxDocs=44421)
                0.01734314 = queryNorm
              0.2893684 = fieldWeight in 4014, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6298943 = idf(docFreq=1177, maxDocs=44421)
                0.0625 = fieldNorm(doc=4014)
          0.1529717 = weight(abstract_txt:distribution in 4014) [ClassicSimilarity], result of:
            0.1529717 = score(doc=4014,freq=2.0), product of:
              0.30923316 = queryWeight, product of:
                3.185872 = boost
                5.5966744 = idf(docFreq=447, maxDocs=44421)
                0.01734314 = queryNorm
              0.4946808 = fieldWeight in 4014, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.5966744 = idf(docFreq=447, maxDocs=44421)
                0.0625 = fieldNorm(doc=4014)
        0.36 = coord(9/25)
    
  5. Spitkovsky, V.; Norvig, P.: From words to concepts and back : dictionaries for linking text, entities and ideas (2012) 0.18
    0.18486579 = sum of:
      0.18486579 = product of:
        0.5777056 = sum of:
          0.030574823 = weight(abstract_txt:areas in 1337) [ClassicSimilarity], result of:
            0.030574823 = score(doc=1337,freq=1.0), product of:
              0.13388886 = queryWeight, product of:
                1.5846686 = boost
                4.871674 = idf(docFreq=924, maxDocs=44421)
                0.01734314 = queryNorm
              0.22835973 = fieldWeight in 1337, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.871674 = idf(docFreq=924, maxDocs=44421)
                0.046875 = fieldNorm(doc=1337)
          0.048722632 = weight(abstract_txt:natural in 1337) [ClassicSimilarity], result of:
            0.048722632 = score(doc=1337,freq=2.0), product of:
              0.14498141 = queryWeight, product of:
                1.6490067 = boost
                5.0694656 = idf(docFreq=758, maxDocs=44421)
                0.01734314 = queryNorm
              0.33606124 = fieldWeight in 1337, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.0694656 = idf(docFreq=758, maxDocs=44421)
                0.046875 = fieldNorm(doc=1337)
          0.03997308 = weight(abstract_txt:each in 1337) [ClassicSimilarity], result of:
            0.03997308 = score(doc=1337,freq=3.0), product of:
              0.11956658 = queryWeight, product of:
                1.6742727 = boost
                4.1177115 = idf(docFreq=1965, maxDocs=44421)
                0.01734314 = queryNorm
              0.3343165 = fieldWeight in 1337, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                4.1177115 = idf(docFreq=1965, maxDocs=44421)
                0.046875 = fieldNorm(doc=1337)
          0.022828234 = weight(abstract_txt:were in 1337) [ClassicSimilarity], result of:
            0.022828234 = score(doc=1337,freq=1.0), product of:
              0.13278918 = queryWeight, product of:
                2.087693 = boost
                3.6674848 = idf(docFreq=3083, maxDocs=44421)
                0.01734314 = queryNorm
              0.17191336 = fieldWeight in 1337, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.6674848 = idf(docFreq=3083, maxDocs=44421)
                0.046875 = fieldNorm(doc=1337)
          0.068278015 = weight(abstract_txt:text in 1337) [ClassicSimilarity], result of:
            0.068278015 = score(doc=1337,freq=5.0), product of:
              0.16120495 = queryWeight, product of:
                2.300247 = boost
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.01734314 = queryNorm
              0.42354786 = fieldWeight in 1337, product of:
                2.236068 = tf(freq=5.0), with freq of:
                  5.0 = termFreq=5.0
                4.040882 = idf(docFreq=2122, maxDocs=44421)
                0.046875 = fieldNorm(doc=1337)
          0.0524895 = weight(abstract_txt:relevant in 1337) [ClassicSimilarity], result of:
            0.0524895 = score(doc=1337,freq=1.0), product of:
              0.2418578 = queryWeight, product of:
                3.012044 = boost
                4.6298943 = idf(docFreq=1177, maxDocs=44421)
                0.01734314 = queryNorm
              0.2170263 = fieldWeight in 1337, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6298943 = idf(docFreq=1177, maxDocs=44421)
                0.046875 = fieldNorm(doc=1337)
          0.12879194 = weight(abstract_txt:phrases in 1337) [ClassicSimilarity], result of:
            0.12879194 = score(doc=1337,freq=1.0), product of:
              0.39975265 = queryWeight, product of:
                3.3535714 = boost
                6.8731537 = idf(docFreq=124, maxDocs=44421)
                0.01734314 = queryNorm
              0.32217908 = fieldWeight in 1337, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.8731537 = idf(docFreq=124, maxDocs=44421)
                0.046875 = fieldNorm(doc=1337)
          0.18604736 = weight(abstract_txt:noun in 1337) [ClassicSimilarity], result of:
            0.18604736 = score(doc=1337,freq=1.0), product of:
              0.5108357 = queryWeight, product of:
                3.790989 = boost
                7.769642 = idf(docFreq=50, maxDocs=44421)
                0.01734314 = queryNorm
              0.36420196 = fieldWeight in 1337, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.769642 = idf(docFreq=50, maxDocs=44421)
                0.046875 = fieldNorm(doc=1337)
        0.32 = coord(8/25)