Document (#40224)

Author
Järvelin, A.
Keskustalo, H.
Sormunen, E.
Saastamoinen, M.
Kettunen, K.
Title
Information retrieval from historical newspaper collections in highly inflectional languages : a query expansion approach
Source
Journal of the Association for Information Science and Technology. 67(2016) no.12, S.2928-2946
Year
2016
Abstract
The aim of the study was to test whether query expansion by approximate string matching methods is beneficial in retrieval from historical newspaper collections in a language rich with compounds and inflectional forms (Finnish). First, approximate string matching methods were used to generate lists of index words most similar to contemporary query terms in a digitized newspaper collection from the 1800s. Top index word variants were categorized to estimate the appropriate query expansion ranges in the retrieval test. Second, the effectiveness of approximate string matching methods, automatically generated inflectional forms, and their combinations were measured in a Cranfield-style test. Finally, a detailed topic-level analysis of test results was conducted. In the index of historical newspaper collection the occurrences of a word typically spread to many linguistic and historical variants along with optical character recognition (OCR) errors. All query expansion methods improved the baseline results. Extensive expansion of around 30 variants for each query word was required to achieve the highest performance improvement. Query expansion based on approximate string matching was superior to using the inflectional forms of the query words, showing that coverage of the different types of variation is more important than precision in handling one type of variation.
Content
Vgl.: http://onlinelibrary.wiley.com/doi/10.1002/asi.23379/full.
Theme
Computerlinguistik
Semantisches Umfeld in Indexierung u. Retrieval
Form
Zeitungen

Similar documents (author)

  1. Järvelin, K.; Kristensen, J.; Niemi, T.; Sormunen, E.; Keskustalo, H.: ¬A deductive data model for query expansion (1996) 4.84
    4.8363953 = sum of:
      4.8363953 = sum of:
        1.1341268 = weight(author_txt:järvelin in 3230) [ClassicSimilarity], result of:
          1.1341268 = score(doc=3230,freq=1.0), product of:
            0.4543382 = queryWeight, product of:
              7.9878955 = idf(docFreq=40, maxDocs=44421)
              0.056878336 = queryNorm
            2.4962173 = fieldWeight in 3230, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              7.9878955 = idf(docFreq=40, maxDocs=44421)
              0.3125 = fieldNorm(doc=3230)
        1.7919003 = weight(author_txt:sormunen in 3230) [ClassicSimilarity], result of:
          1.7919003 = score(doc=3230,freq=1.0), product of:
            0.61633104 = queryWeight, product of:
              1.1647089 = boost
              9.303573 = idf(docFreq=10, maxDocs=44421)
              0.056878336 = queryNorm
            2.9073665 = fieldWeight in 3230, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              9.303573 = idf(docFreq=10, maxDocs=44421)
              0.3125 = fieldNorm(doc=3230)
        1.9103683 = weight(author_txt:keskustalo in 3230) [ClassicSimilarity], result of:
          1.9103683 = score(doc=3230,freq=1.0), product of:
            0.6432052 = queryWeight, product of:
              1.1898307 = boost
              9.504243 = idf(docFreq=8, maxDocs=44421)
              0.056878336 = queryNorm
            2.9700758 = fieldWeight in 3230, product of:
              1.0 = tf(freq=1.0), with freq of:
                1.0 = termFreq=1.0
              9.504243 = idf(docFreq=8, maxDocs=44421)
              0.3125 = fieldNorm(doc=3230)
    
  2. Lehtokangas, R.; Keskustalo, H.; Järvelin, K.: Experiments with transitive dictionary translation and pseudo-relevance feedback using graded relevance assessments (2008) 2.44
    2.4355962 = sum of:
      2.4355962 = product of:
        3.6533942 = sum of:
          1.3609523 = weight(author_txt:järvelin in 2349) [ClassicSimilarity], result of:
            1.3609523 = score(doc=2349,freq=1.0), product of:
              0.4543382 = queryWeight, product of:
                7.9878955 = idf(docFreq=40, maxDocs=44421)
                0.056878336 = queryNorm
              2.9954607 = fieldWeight in 2349, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.9878955 = idf(docFreq=40, maxDocs=44421)
                0.375 = fieldNorm(doc=2349)
          2.292442 = weight(author_txt:keskustalo in 2349) [ClassicSimilarity], result of:
            2.292442 = score(doc=2349,freq=1.0), product of:
              0.6432052 = queryWeight, product of:
                1.1898307 = boost
                9.504243 = idf(docFreq=8, maxDocs=44421)
                0.056878336 = queryNorm
              3.5640912 = fieldWeight in 2349, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.504243 = idf(docFreq=8, maxDocs=44421)
                0.375 = fieldNorm(doc=2349)
        0.6666667 = coord(2/3)
    
  3. Pirkola, A.; Hedlund, T.; Keskustalo, H.; Järvelin, K.: Dictionary-based cross-language information retrieval : problems, methods, and research findings (2001) 2.03
    2.0296636 = sum of:
      2.0296636 = product of:
        3.044495 = sum of:
          1.1341268 = weight(author_txt:järvelin in 4908) [ClassicSimilarity], result of:
            1.1341268 = score(doc=4908,freq=1.0), product of:
              0.4543382 = queryWeight, product of:
                7.9878955 = idf(docFreq=40, maxDocs=44421)
                0.056878336 = queryNorm
              2.4962173 = fieldWeight in 4908, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.9878955 = idf(docFreq=40, maxDocs=44421)
                0.3125 = fieldNorm(doc=4908)
          1.9103683 = weight(author_txt:keskustalo in 4908) [ClassicSimilarity], result of:
            1.9103683 = score(doc=4908,freq=1.0), product of:
              0.6432052 = queryWeight, product of:
                1.1898307 = boost
                9.504243 = idf(docFreq=8, maxDocs=44421)
                0.056878336 = queryNorm
              2.9700758 = fieldWeight in 4908, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.504243 = idf(docFreq=8, maxDocs=44421)
                0.3125 = fieldNorm(doc=4908)
        0.6666667 = coord(2/3)
    
  4. Toivonen, J.; Pirkola, A.; Keskustalo, H.; Visala, K.; Järvelin, K.: Translating cross-lingual spelling variants using transformation rules (2005) 2.03
    2.0296636 = sum of:
      2.0296636 = product of:
        3.044495 = sum of:
          1.1341268 = weight(author_txt:järvelin in 2052) [ClassicSimilarity], result of:
            1.1341268 = score(doc=2052,freq=1.0), product of:
              0.4543382 = queryWeight, product of:
                7.9878955 = idf(docFreq=40, maxDocs=44421)
                0.056878336 = queryNorm
              2.4962173 = fieldWeight in 2052, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.9878955 = idf(docFreq=40, maxDocs=44421)
                0.3125 = fieldNorm(doc=2052)
          1.9103683 = weight(author_txt:keskustalo in 2052) [ClassicSimilarity], result of:
            1.9103683 = score(doc=2052,freq=1.0), product of:
              0.6432052 = queryWeight, product of:
                1.1898307 = boost
                9.504243 = idf(docFreq=8, maxDocs=44421)
                0.056878336 = queryNorm
              2.9700758 = fieldWeight in 2052, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.504243 = idf(docFreq=8, maxDocs=44421)
                0.3125 = fieldNorm(doc=2052)
        0.6666667 = coord(2/3)
    
  5. Ferro, N.; Silvello, G.; Keskustalo, H.; Pirkola, A.; Järvelin, K.: ¬The twist measure for IR evaluation : taking user's effort into account (2016) 2.03
    2.0296636 = sum of:
      2.0296636 = product of:
        3.044495 = sum of:
          1.1341268 = weight(author_txt:järvelin in 3771) [ClassicSimilarity], result of:
            1.1341268 = score(doc=3771,freq=1.0), product of:
              0.4543382 = queryWeight, product of:
                7.9878955 = idf(docFreq=40, maxDocs=44421)
                0.056878336 = queryNorm
              2.4962173 = fieldWeight in 3771, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.9878955 = idf(docFreq=40, maxDocs=44421)
                0.3125 = fieldNorm(doc=3771)
          1.9103683 = weight(author_txt:keskustalo in 3771) [ClassicSimilarity], result of:
            1.9103683 = score(doc=3771,freq=1.0), product of:
              0.6432052 = queryWeight, product of:
                1.1898307 = boost
                9.504243 = idf(docFreq=8, maxDocs=44421)
                0.056878336 = queryNorm
              2.9700758 = fieldWeight in 3771, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.504243 = idf(docFreq=8, maxDocs=44421)
                0.3125 = fieldNorm(doc=3771)
        0.6666667 = coord(2/3)
    

Similar documents (content)

  1. French, J.C.; Powell, A.L.; Schulman, E.: Using clustering strategies for creating authority files (2000) 0.33
    0.33470434 = sum of:
      0.33470434 = product of:
        1.0459511 = sum of:
          0.00810409 = weight(abstract_txt:from in 5811) [ClassicSimilarity], result of:
            0.00810409 = score(doc=5811,freq=1.0), product of:
              0.03759237 = queryWeight, product of:
                1.1149384 = boost
                2.759399 = idf(docFreq=7646, maxDocs=44421)
                0.012218963 = queryNorm
              0.21557805 = fieldWeight in 5811, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                2.759399 = idf(docFreq=7646, maxDocs=44421)
                0.078125 = fieldNorm(doc=5811)
          0.016206436 = weight(abstract_txt:retrieval in 5811) [ClassicSimilarity], result of:
            0.016206436 = score(doc=5811,freq=1.0), product of:
              0.059669893 = queryWeight, product of:
                1.4046841 = boost
                3.4765 = idf(docFreq=3732, maxDocs=44421)
                0.012218963 = queryNorm
              0.27160156 = fieldWeight in 5811, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.4765 = idf(docFreq=3732, maxDocs=44421)
                0.078125 = fieldNorm(doc=5811)
          0.062029038 = weight(abstract_txt:word in 5811) [ClassicSimilarity], result of:
            0.062029038 = score(doc=5811,freq=1.0), product of:
              0.1460025 = queryWeight, product of:
                2.1972587 = boost
                5.4380693 = idf(docFreq=524, maxDocs=44421)
                0.012218963 = queryNorm
              0.42484915 = fieldWeight in 5811, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.4380693 = idf(docFreq=524, maxDocs=44421)
                0.078125 = fieldNorm(doc=5811)
          0.08846752 = weight(abstract_txt:forms in 5811) [ClassicSimilarity], result of:
            0.08846752 = score(doc=5811,freq=2.0), product of:
              0.1468282 = queryWeight, product of:
                2.203463 = boost
                5.453425 = idf(docFreq=516, maxDocs=44421)
                0.012218963 = queryNorm
              0.60252404 = fieldWeight in 5811, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.453425 = idf(docFreq=516, maxDocs=44421)
                0.078125 = fieldNorm(doc=5811)
          0.16251156 = weight(abstract_txt:variants in 5811) [ClassicSimilarity], result of:
            0.16251156 = score(doc=5811,freq=1.0), product of:
              0.27747238 = queryWeight, product of:
                3.029081 = boost
                7.496775 = idf(docFreq=66, maxDocs=44421)
                0.012218963 = queryNorm
              0.58568555 = fieldWeight in 5811, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.496775 = idf(docFreq=66, maxDocs=44421)
                0.078125 = fieldNorm(doc=5811)
          0.16041817 = weight(abstract_txt:matching in 5811) [ClassicSimilarity], result of:
            0.16041817 = score(doc=5811,freq=2.0), product of:
              0.24030834 = queryWeight, product of:
                3.2550287 = boost
                6.0419855 = idf(docFreq=286, maxDocs=44421)
                0.012218963 = queryNorm
              0.6675514 = fieldWeight in 5811, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                6.0419855 = idf(docFreq=286, maxDocs=44421)
                0.078125 = fieldNorm(doc=5811)
          0.19033323 = weight(abstract_txt:string in 5811) [ClassicSimilarity], result of:
            0.19033323 = score(doc=5811,freq=1.0), product of:
              0.3393279 = queryWeight, product of:
                3.867944 = boost
                7.179679 = idf(docFreq=91, maxDocs=44421)
                0.012218963 = queryNorm
              0.56091243 = fieldWeight in 5811, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.179679 = idf(docFreq=91, maxDocs=44421)
                0.078125 = fieldNorm(doc=5811)
          0.35788107 = weight(abstract_txt:approximate in 5811) [ClassicSimilarity], result of:
            0.35788107 = score(doc=5811,freq=2.0), product of:
              0.41029128 = queryWeight, product of:
                4.253207 = boost
                7.894805 = idf(docFreq=44, maxDocs=44421)
                0.012218963 = queryNorm
              0.8722609 = fieldWeight in 5811, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                7.894805 = idf(docFreq=44, maxDocs=44421)
                0.078125 = fieldNorm(doc=5811)
        0.32 = coord(8/25)
    
  2. Galvez, C.; Moya-Anegón, F.: Approximate personal name-matching through finite-state graphs (2007) 0.33
    0.3304485 = sum of:
      0.3304485 = product of:
        0.9179125 = sum of:
          0.011229356 = weight(abstract_txt:from in 1614) [ClassicSimilarity], result of:
            0.011229356 = score(doc=1614,freq=3.0), product of:
              0.03759237 = queryWeight, product of:
                1.1149384 = boost
                2.759399 = idf(docFreq=7646, maxDocs=44421)
                0.012218963 = queryNorm
              0.29871368 = fieldWeight in 1614, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                2.759399 = idf(docFreq=7646, maxDocs=44421)
                0.0625 = fieldNorm(doc=1614)
          0.012965149 = weight(abstract_txt:retrieval in 1614) [ClassicSimilarity], result of:
            0.012965149 = score(doc=1614,freq=1.0), product of:
              0.059669893 = queryWeight, product of:
                1.4046841 = boost
                3.4765 = idf(docFreq=3732, maxDocs=44421)
                0.012218963 = queryNorm
              0.21728125 = fieldWeight in 1614, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.4765 = idf(docFreq=3732, maxDocs=44421)
                0.0625 = fieldNorm(doc=1614)
          0.046758767 = weight(abstract_txt:index in 1614) [ClassicSimilarity], result of:
            0.046758767 = score(doc=1614,freq=2.0), product of:
              0.11137874 = queryWeight, product of:
                1.9191202 = boost
                4.7496953 = idf(docFreq=1044, maxDocs=44421)
                0.012218963 = queryNorm
              0.41981772 = fieldWeight in 1614, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.7496953 = idf(docFreq=1044, maxDocs=44421)
                0.0625 = fieldNorm(doc=1614)
          0.10008957 = weight(abstract_txt:forms in 1614) [ClassicSimilarity], result of:
            0.10008957 = score(doc=1614,freq=4.0), product of:
              0.1468282 = queryWeight, product of:
                2.203463 = boost
                5.453425 = idf(docFreq=516, maxDocs=44421)
                0.012218963 = queryNorm
              0.6816781 = fieldWeight in 1614, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                5.453425 = idf(docFreq=516, maxDocs=44421)
                0.0625 = fieldNorm(doc=1614)
          0.04139024 = weight(abstract_txt:methods in 1614) [ClassicSimilarity], result of:
            0.04139024 = score(doc=1614,freq=2.0), product of:
              0.11301562 = queryWeight, product of:
                2.2322335 = boost
                4.1434727 = idf(docFreq=1915, maxDocs=44421)
                0.012218963 = queryNorm
              0.3662347 = fieldWeight in 1614, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.1434727 = idf(docFreq=1915, maxDocs=44421)
                0.0625 = fieldNorm(doc=1614)
          0.2600185 = weight(abstract_txt:variants in 1614) [ClassicSimilarity], result of:
            0.2600185 = score(doc=1614,freq=4.0), product of:
              0.27747238 = queryWeight, product of:
                3.029081 = boost
                7.496775 = idf(docFreq=66, maxDocs=44421)
                0.012218963 = queryNorm
              0.9370969 = fieldWeight in 1614, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                7.496775 = idf(docFreq=66, maxDocs=44421)
                0.0625 = fieldNorm(doc=1614)
          0.090746224 = weight(abstract_txt:matching in 1614) [ClassicSimilarity], result of:
            0.090746224 = score(doc=1614,freq=1.0), product of:
              0.24030834 = queryWeight, product of:
                3.2550287 = boost
                6.0419855 = idf(docFreq=286, maxDocs=44421)
                0.012218963 = queryNorm
              0.3776241 = fieldWeight in 1614, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.0419855 = idf(docFreq=286, maxDocs=44421)
                0.0625 = fieldNorm(doc=1614)
          0.15226659 = weight(abstract_txt:string in 1614) [ClassicSimilarity], result of:
            0.15226659 = score(doc=1614,freq=1.0), product of:
              0.3393279 = queryWeight, product of:
                3.867944 = boost
                7.179679 = idf(docFreq=91, maxDocs=44421)
                0.012218963 = queryNorm
              0.44872993 = fieldWeight in 1614, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.179679 = idf(docFreq=91, maxDocs=44421)
                0.0625 = fieldNorm(doc=1614)
          0.2024481 = weight(abstract_txt:approximate in 1614) [ClassicSimilarity], result of:
            0.2024481 = score(doc=1614,freq=1.0), product of:
              0.41029128 = queryWeight, product of:
                4.253207 = boost
                7.894805 = idf(docFreq=44, maxDocs=44421)
                0.012218963 = queryNorm
              0.4934253 = fieldWeight in 1614, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.894805 = idf(docFreq=44, maxDocs=44421)
                0.0625 = fieldNorm(doc=1614)
        0.36 = coord(9/25)
    
  3. Pirkola, A.; Puolamäki, D.; Järvelin, K.: Applying query structuring in cross-language retrieval (2003) 0.30
    0.30234677 = sum of:
      0.30234677 = product of:
        0.94483364 = sum of:
          0.10444873 = weight(abstract_txt:finnish in 2074) [ClassicSimilarity], result of:
            0.10444873 = score(doc=2074,freq=3.0), product of:
              0.09934626 = queryWeight, product of:
                1.0464444 = boost
                7.769642 = idf(docFreq=50, maxDocs=44421)
                0.012218963 = queryNorm
              1.0513605 = fieldWeight in 2074, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                7.769642 = idf(docFreq=50, maxDocs=44421)
                0.078125 = fieldNorm(doc=2074)
          0.02807037 = weight(abstract_txt:retrieval in 2074) [ClassicSimilarity], result of:
            0.02807037 = score(doc=2074,freq=3.0), product of:
              0.059669893 = queryWeight, product of:
                1.4046841 = boost
                3.4765 = idf(docFreq=3732, maxDocs=44421)
                0.012218963 = queryNorm
              0.4704277 = fieldWeight in 2074, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                3.4765 = idf(docFreq=3732, maxDocs=44421)
                0.078125 = fieldNorm(doc=2074)
          0.038053602 = weight(abstract_txt:were in 2074) [ClassicSimilarity], result of:
            0.038053602 = score(doc=2074,freq=4.0), product of:
              0.06640602 = queryWeight, product of:
                1.4818517 = boost
                3.6674848 = idf(docFreq=3083, maxDocs=44421)
                0.012218963 = queryNorm
              0.5730445 = fieldWeight in 2074, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                3.6674848 = idf(docFreq=3083, maxDocs=44421)
                0.078125 = fieldNorm(doc=2074)
          0.09344656 = weight(abstract_txt:test in 2074) [ClassicSimilarity], result of:
            0.09344656 = score(doc=2074,freq=2.0), product of:
              0.16761339 = queryWeight, product of:
                2.7184713 = boost
                5.046027 = idf(docFreq=776, maxDocs=44421)
                0.012218963 = queryNorm
              0.5575125 = fieldWeight in 2074, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                5.046027 = idf(docFreq=776, maxDocs=44421)
                0.078125 = fieldNorm(doc=2074)
          0.16251156 = weight(abstract_txt:variants in 2074) [ClassicSimilarity], result of:
            0.16251156 = score(doc=2074,freq=1.0), product of:
              0.27747238 = queryWeight, product of:
                3.029081 = boost
                7.496775 = idf(docFreq=66, maxDocs=44421)
                0.012218963 = queryNorm
              0.58568555 = fieldWeight in 2074, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.496775 = idf(docFreq=66, maxDocs=44421)
                0.078125 = fieldNorm(doc=2074)
          0.11343277 = weight(abstract_txt:matching in 2074) [ClassicSimilarity], result of:
            0.11343277 = score(doc=2074,freq=1.0), product of:
              0.24030834 = queryWeight, product of:
                3.2550287 = boost
                6.0419855 = idf(docFreq=286, maxDocs=44421)
                0.012218963 = queryNorm
              0.4720301 = fieldWeight in 2074, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.0419855 = idf(docFreq=286, maxDocs=44421)
                0.078125 = fieldNorm(doc=2074)
          0.18377861 = weight(abstract_txt:newspaper in 2074) [ClassicSimilarity], result of:
            0.18377861 = score(doc=2074,freq=1.0), product of:
              0.33149207 = queryWeight, product of:
                3.8230233 = boost
                7.0962973 = idf(docFreq=99, maxDocs=44421)
                0.012218963 = queryNorm
              0.55439824 = fieldWeight in 2074, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.0962973 = idf(docFreq=99, maxDocs=44421)
                0.078125 = fieldNorm(doc=2074)
          0.2210914 = weight(abstract_txt:query in 2074) [ClassicSimilarity], result of:
            0.2210914 = score(doc=2074,freq=4.0), product of:
              0.29761013 = queryWeight, product of:
                5.122822 = boost
                4.754492 = idf(docFreq=1039, maxDocs=44421)
                0.012218963 = queryNorm
              0.74288934 = fieldWeight in 2074, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                4.754492 = idf(docFreq=1039, maxDocs=44421)
                0.078125 = fieldNorm(doc=2074)
        0.32 = coord(8/25)
    
  4. Bellaachia, A.; Amor-Tijani, G.: Proper nouns in English-Arabic cross language information retrieval (2008) 0.29
    0.288629 = sum of:
      0.288629 = product of:
        0.9019656 = sum of:
          0.012965149 = weight(abstract_txt:retrieval in 3372) [ClassicSimilarity], result of:
            0.012965149 = score(doc=3372,freq=1.0), product of:
              0.059669893 = queryWeight, product of:
                1.4046841 = boost
                3.4765 = idf(docFreq=3732, maxDocs=44421)
                0.012218963 = queryNorm
              0.21728125 = fieldWeight in 3372, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.4765 = idf(docFreq=3732, maxDocs=44421)
                0.0625 = fieldNorm(doc=3372)
          0.05473949 = weight(abstract_txt:words in 3372) [ClassicSimilarity], result of:
            0.05473949 = score(doc=3372,freq=3.0), product of:
              0.09441332 = queryWeight, product of:
                1.4426867 = boost
                5.355831 = idf(docFreq=569, maxDocs=44421)
                0.012218963 = queryNorm
              0.5797857 = fieldWeight in 3372, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                5.355831 = idf(docFreq=569, maxDocs=44421)
                0.0625 = fieldNorm(doc=3372)
          0.03306344 = weight(abstract_txt:index in 3372) [ClassicSimilarity], result of:
            0.03306344 = score(doc=3372,freq=1.0), product of:
              0.11137874 = queryWeight, product of:
                1.9191202 = boost
                4.7496953 = idf(docFreq=1044, maxDocs=44421)
                0.012218963 = queryNorm
              0.29685596 = fieldWeight in 3372, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.7496953 = idf(docFreq=1044, maxDocs=44421)
                0.0625 = fieldNorm(doc=3372)
          0.13000925 = weight(abstract_txt:variants in 3372) [ClassicSimilarity], result of:
            0.13000925 = score(doc=3372,freq=1.0), product of:
              0.27747238 = queryWeight, product of:
                3.029081 = boost
                7.496775 = idf(docFreq=66, maxDocs=44421)
                0.012218963 = queryNorm
              0.46854845 = fieldWeight in 3372, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.496775 = idf(docFreq=66, maxDocs=44421)
                0.0625 = fieldNorm(doc=3372)
          0.12833454 = weight(abstract_txt:matching in 3372) [ClassicSimilarity], result of:
            0.12833454 = score(doc=3372,freq=2.0), product of:
              0.24030834 = queryWeight, product of:
                3.2550287 = boost
                6.0419855 = idf(docFreq=286, maxDocs=44421)
                0.012218963 = queryNorm
              0.5340411 = fieldWeight in 3372, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                6.0419855 = idf(docFreq=286, maxDocs=44421)
                0.0625 = fieldNorm(doc=3372)
          0.21533746 = weight(abstract_txt:string in 3372) [ClassicSimilarity], result of:
            0.21533746 = score(doc=3372,freq=2.0), product of:
              0.3393279 = queryWeight, product of:
                3.867944 = boost
                7.179679 = idf(docFreq=91, maxDocs=44421)
                0.012218963 = queryNorm
              0.6345999 = fieldWeight in 3372, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                7.179679 = idf(docFreq=91, maxDocs=44421)
                0.0625 = fieldNorm(doc=3372)
          0.2024481 = weight(abstract_txt:approximate in 3372) [ClassicSimilarity], result of:
            0.2024481 = score(doc=3372,freq=1.0), product of:
              0.41029128 = queryWeight, product of:
                4.253207 = boost
                7.894805 = idf(docFreq=44, maxDocs=44421)
                0.012218963 = queryNorm
              0.4934253 = fieldWeight in 3372, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.894805 = idf(docFreq=44, maxDocs=44421)
                0.0625 = fieldNorm(doc=3372)
          0.12506817 = weight(abstract_txt:query in 3372) [ClassicSimilarity], result of:
            0.12506817 = score(doc=3372,freq=2.0), product of:
              0.29761013 = queryWeight, product of:
                5.122822 = boost
                4.754492 = idf(docFreq=1039, maxDocs=44421)
                0.012218963 = queryNorm
              0.42024165 = fieldWeight in 3372, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                4.754492 = idf(docFreq=1039, maxDocs=44421)
                0.0625 = fieldNorm(doc=3372)
        0.32 = coord(8/25)
    
  5. Airio, E.; Kettunen, K.: Does dictionary based bilingual retrieval work in a non-normalized index? (2009) 0.26
    0.25761378 = sum of:
      0.25761378 = product of:
        0.7155938 = sum of:
          0.09648561 = weight(abstract_txt:finnish in 224) [ClassicSimilarity], result of:
            0.09648561 = score(doc=224,freq=4.0), product of:
              0.09934626 = queryWeight, product of:
                1.0464444 = boost
                7.769642 = idf(docFreq=50, maxDocs=44421)
                0.012218963 = queryNorm
              0.97120523 = fieldWeight in 224, product of:
                2.0 = tf(freq=4.0), with freq of:
                  4.0 = termFreq=4.0
                7.769642 = idf(docFreq=50, maxDocs=44421)
                0.0625 = fieldNorm(doc=224)
          0.022456296 = weight(abstract_txt:retrieval in 224) [ClassicSimilarity], result of:
            0.022456296 = score(doc=224,freq=3.0), product of:
              0.059669893 = queryWeight, product of:
                1.4046841 = boost
                3.4765 = idf(docFreq=3732, maxDocs=44421)
                0.012218963 = queryNorm
              0.37634215 = fieldWeight in 224, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                3.4765 = idf(docFreq=3732, maxDocs=44421)
                0.0625 = fieldNorm(doc=224)
          0.015221441 = weight(abstract_txt:were in 224) [ClassicSimilarity], result of:
            0.015221441 = score(doc=224,freq=1.0), product of:
              0.06640602 = queryWeight, product of:
                1.4818517 = boost
                3.6674848 = idf(docFreq=3083, maxDocs=44421)
                0.012218963 = queryNorm
              0.2292178 = fieldWeight in 224, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.6674848 = idf(docFreq=3083, maxDocs=44421)
                0.0625 = fieldNorm(doc=224)
          0.03306344 = weight(abstract_txt:index in 224) [ClassicSimilarity], result of:
            0.03306344 = score(doc=224,freq=1.0), product of:
              0.11137874 = queryWeight, product of:
                1.9191202 = boost
                4.7496953 = idf(docFreq=1044, maxDocs=44421)
                0.012218963 = queryNorm
              0.29685596 = fieldWeight in 224, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.7496953 = idf(docFreq=1044, maxDocs=44421)
                0.0625 = fieldNorm(doc=224)
          0.050044786 = weight(abstract_txt:forms in 224) [ClassicSimilarity], result of:
            0.050044786 = score(doc=224,freq=1.0), product of:
              0.1468282 = queryWeight, product of:
                2.203463 = boost
                5.453425 = idf(docFreq=516, maxDocs=44421)
                0.012218963 = queryNorm
              0.34083906 = fieldWeight in 224, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.453425 = idf(docFreq=516, maxDocs=44421)
                0.0625 = fieldNorm(doc=224)
          0.052861355 = weight(abstract_txt:test in 224) [ClassicSimilarity], result of:
            0.052861355 = score(doc=224,freq=1.0), product of:
              0.16761339 = queryWeight, product of:
                2.7184713 = boost
                5.046027 = idf(docFreq=776, maxDocs=44421)
                0.012218963 = queryNorm
              0.3153767 = fieldWeight in 224, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.046027 = idf(docFreq=776, maxDocs=44421)
                0.0625 = fieldNorm(doc=224)
          0.090746224 = weight(abstract_txt:matching in 224) [ClassicSimilarity], result of:
            0.090746224 = score(doc=224,freq=1.0), product of:
              0.24030834 = queryWeight, product of:
                3.2550287 = boost
                6.0419855 = idf(docFreq=286, maxDocs=44421)
                0.012218963 = queryNorm
              0.3776241 = fieldWeight in 224, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.0419855 = idf(docFreq=286, maxDocs=44421)
                0.0625 = fieldNorm(doc=224)
          0.15226659 = weight(abstract_txt:string in 224) [ClassicSimilarity], result of:
            0.15226659 = score(doc=224,freq=1.0), product of:
              0.3393279 = queryWeight, product of:
                3.867944 = boost
                7.179679 = idf(docFreq=91, maxDocs=44421)
                0.012218963 = queryNorm
              0.44872993 = fieldWeight in 224, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.179679 = idf(docFreq=91, maxDocs=44421)
                0.0625 = fieldNorm(doc=224)
          0.2024481 = weight(abstract_txt:approximate in 224) [ClassicSimilarity], result of:
            0.2024481 = score(doc=224,freq=1.0), product of:
              0.41029128 = queryWeight, product of:
                4.253207 = boost
                7.894805 = idf(docFreq=44, maxDocs=44421)
                0.012218963 = queryNorm
              0.4934253 = fieldWeight in 224, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.894805 = idf(docFreq=44, maxDocs=44421)
                0.0625 = fieldNorm(doc=224)
        0.36 = coord(9/25)