Document (#32320)

Author
Kaufmann, E.
Title
¬Das Indexieren von natürlichsprachlichen Dokumenten und die inverse Seitenhäufigkeit
Source
http://www.ifi.unizh.ch/cl/study/lizarbeiten/lizkaufmann.pdf
Year
2001
Abstract
Die Lizentiatsarbeit gibt im ersten theoretischen Teil einen Überblick über das Indexieren von Dokumenten. Sie zeigt die verschiedenen Typen von Indexen sowie die wichtigsten Aspekte bezüglich einer Indexsprache auf. Diverse manuelle und automatische Indexierungsverfahren werden präsentiert. Spezielle Aufmerksamkeit innerhalb des ersten Teils gilt den Schlagwortregistern, deren charakteristische Merkmale und Eigenheiten erörtert werden. Zusätzlich werden die gängigen Kriterien zur Bewertung von Indexen sowie die Masse zur Evaluation von Indexierungsverfahren und Indexierungsergebnissen vorgestellt. Im zweiten Teil der Arbeit werden fünf reale Bücher einer statistischen Untersuchung unterzogen. Zum einen werden die lexikalischen und syntaktischen Bestandteile der fünf Buchregister ermittelt, um den Inhalt von Schlagwortregistern zu erschliessen. Andererseits werden aus den Textausschnitten der Bücher Indexterme maschinell extrahiert und mit den Schlagworteinträgen in den Buchregistern verglichen. Das Hauptziel der Untersuchungen besteht darin, eine Indexierungsmethode, die auf linguistikorientierter Extraktion der Indexterme und Termhäufigkeitsgewichtung basiert, im Hinblick auf ihren Gebrauchswert für eine automatische Indexierung zu testen. Die Gewichtungsmethode ist die inverse Seitenhäufigkeit, eine Methode, welche von der inversen Dokumentfrequenz abgeleitet wurde, zur automatischen Erstellung von Schlagwortregistern für deutschsprachige Texte. Die Prüfung der Methode im statistischen Teil führte nicht zu zufriedenstellenden Resultaten.
Content
Lizentiatsarbeit der Philosphischen Fakultät der Universität Zürich, - Vgl. auch: http://www.ifi.unizh.ch/cl/study/lizarbeiten/lizkaufmann.pdf.
Theme
Automatisches Indexieren
Register

Similar documents (author)

  1. Kaufmann, N.C.: Kommt das Domainsterben? : Rechtsprechung gibt beschreibende Internet-Adressen zum Abschuss frei (2000) 5.99
    5.989656 = sum of:
      5.989656 = weight(author_txt:kaufmann in 5279) [ClassicSimilarity], result of:
        5.989656 = fieldWeight in 5279, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          9.583449 = idf(docFreq=7, maxDocs=42740)
          0.625 = fieldNorm(doc=5279)
    
  2. Kaufmann, T.: Googeln wie die Profis : Perfekte Suche (2004) 5.99
    5.989656 = sum of:
      5.989656 = weight(author_txt:kaufmann in 2927) [ClassicSimilarity], result of:
        5.989656 = fieldWeight in 2927, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          9.583449 = idf(docFreq=7, maxDocs=42740)
          0.625 = fieldNorm(doc=2927)
    
  3. Havemann, F.; Kaufmann, A.: ¬Der Wandel des Benutzerverhaltens in Zeiten des Internet : Ergebnisse von Befragungen an 13 Bibliotheken (2006) 4.79
    4.7917247 = sum of:
      4.7917247 = weight(author_txt:kaufmann in 1150) [ClassicSimilarity], result of:
        4.7917247 = fieldWeight in 1150, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          9.583449 = idf(docFreq=7, maxDocs=42740)
          0.5 = fieldNorm(doc=1150)
    
  4. Kaufmann, J.-C.: Wenn ICH ein anderer ist (2010) 4.79
    4.7917247 = sum of:
      4.7917247 = weight(author_txt:kaufmann in 639) [ClassicSimilarity], result of:
        4.7917247 = fieldWeight in 639, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          9.583449 = idf(docFreq=7, maxDocs=42740)
          0.5 = fieldNorm(doc=639)
    
  5. Kaufmann, J.-C.: ¬Die Erfindung des Ich : eine Theorie der Identität (2005) 4.79
    4.7917247 = sum of:
      4.7917247 = weight(author_txt:kaufmann in 640) [ClassicSimilarity], result of:
        4.7917247 = fieldWeight in 640, product of:
          1.0 = tf(freq=1.0), with freq of:
            1.0 = termFreq=1.0
          9.583449 = idf(docFreq=7, maxDocs=42740)
          0.5 = fieldNorm(doc=640)
    

Similar documents (content)

  1. Halip, I.: Automatische Extrahierung von Schlagworten aus unstrukturierten Texten (2005) 0.19
    0.19071232 = sum of:
      0.19071232 = product of:
        0.595976 = sum of:
          0.07633294 = weight(abstract_txt:manuelle in 1987) [ClassicSimilarity], result of:
            0.07633294 = score(doc=1987,freq=1.0), product of:
              0.15808079 = queryWeight, product of:
                1.0065156 = boost
                8.829678 = idf(docFreq=16, maxDocs=42740)
                0.017787453 = queryNorm
              0.482873 = fieldWeight in 1987, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                8.829678 = idf(docFreq=16, maxDocs=42740)
                0.0546875 = fieldNorm(doc=1987)
          0.022193167 = weight(abstract_txt:sowie in 1987) [ClassicSimilarity], result of:
            0.022193167 = score(doc=1987,freq=1.0), product of:
              0.0874099 = queryWeight, product of:
                1.0584644 = boost
                4.6426997 = idf(docFreq=1118, maxDocs=42740)
                0.017787453 = queryNorm
              0.25389764 = fieldWeight in 1987, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                4.6426997 = idf(docFreq=1118, maxDocs=42740)
                0.0546875 = fieldNorm(doc=1987)
          0.02064537 = weight(abstract_txt:eine in 1987) [ClassicSimilarity], result of:
            0.02064537 = score(doc=1987,freq=2.0), product of:
              0.07568038 = queryWeight, product of:
                1.2062386 = boost
                3.5272505 = idf(docFreq=3413, maxDocs=42740)
                0.017787453 = queryNorm
              0.27279687 = fieldWeight in 1987, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.5272505 = idf(docFreq=3413, maxDocs=42740)
                0.0546875 = fieldNorm(doc=1987)
          0.10350003 = weight(abstract_txt:dokumenten in 1987) [ClassicSimilarity], result of:
            0.10350003 = score(doc=1987,freq=3.0), product of:
              0.16917428 = queryWeight, product of:
                1.4725266 = boost
                6.458884 = idf(docFreq=181, maxDocs=42740)
                0.017787453 = queryNorm
              0.6117953 = fieldWeight in 1987, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                6.458884 = idf(docFreq=181, maxDocs=42740)
                0.0546875 = fieldNorm(doc=1987)
          0.07398459 = weight(abstract_txt:automatische in 1987) [ClassicSimilarity], result of:
            0.07398459 = score(doc=1987,freq=1.0), product of:
              0.19506317 = queryWeight, product of:
                1.5811883 = boost
                6.9355025 = idf(docFreq=112, maxDocs=42740)
                0.017787453 = queryNorm
              0.3792853 = fieldWeight in 1987, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.9355025 = idf(docFreq=112, maxDocs=42740)
                0.0546875 = fieldNorm(doc=1987)
          0.056507524 = weight(abstract_txt:teil in 1987) [ClassicSimilarity], result of:
            0.056507524 = score(doc=1987,freq=1.0), product of:
              0.1865731 = queryWeight, product of:
                1.8939395 = boost
                5.538207 = idf(docFreq=456, maxDocs=42740)
                0.017787453 = queryNorm
              0.3028707 = fieldWeight in 1987, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.538207 = idf(docFreq=456, maxDocs=42740)
                0.0546875 = fieldNorm(doc=1987)
          0.1714547 = weight(abstract_txt:indexierungsverfahren in 1987) [ClassicSimilarity], result of:
            0.1714547 = score(doc=1987,freq=1.0), product of:
              0.341597 = queryWeight, product of:
                2.09244 = boost
                9.177984 = idf(docFreq=11, maxDocs=42740)
                0.017787453 = queryNorm
              0.501921 = fieldWeight in 1987, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.177984 = idf(docFreq=11, maxDocs=42740)
                0.0546875 = fieldNorm(doc=1987)
          0.071357645 = weight(abstract_txt:werden in 1987) [ClassicSimilarity], result of:
            0.071357645 = score(doc=1987,freq=6.0), product of:
              0.15113491 = queryWeight, product of:
                2.4106767 = boost
                3.524618 = idf(docFreq=3422, maxDocs=42740)
                0.017787453 = queryNorm
              0.47214538 = fieldWeight in 1987, product of:
                2.4494898 = tf(freq=6.0), with freq of:
                  6.0 = termFreq=6.0
                3.524618 = idf(docFreq=3422, maxDocs=42740)
                0.0546875 = fieldNorm(doc=1987)
        0.32 = coord(8/25)
    
  2. Bredack, J.: Automatische Extraktion fachterminologischer Mehrwortbegriffe : ein Verfahrensvergleich (2016) 0.15
    0.15310831 = sum of:
      0.15310831 = product of:
        0.6379513 = sum of:
          0.095433064 = weight(abstract_txt:extrahiert in 5195) [ClassicSimilarity], result of:
            0.095433064 = score(doc=5195,freq=1.0), product of:
              0.16783236 = queryWeight, product of:
                1.0370957 = boost
                9.097941 = idf(docFreq=12, maxDocs=42740)
                0.017787453 = queryNorm
              0.56862134 = fieldWeight in 5195, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.097941 = idf(docFreq=12, maxDocs=42740)
                0.0625 = fieldNorm(doc=5195)
          0.02359471 = weight(abstract_txt:eine in 5195) [ClassicSimilarity], result of:
            0.02359471 = score(doc=5195,freq=2.0), product of:
              0.07568038 = queryWeight, product of:
                1.2062386 = boost
                3.5272505 = idf(docFreq=3413, maxDocs=42740)
                0.017787453 = queryNorm
              0.31176785 = fieldWeight in 5195, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.5272505 = idf(docFreq=3413, maxDocs=42740)
                0.0625 = fieldNorm(doc=5195)
          0.08455382 = weight(abstract_txt:automatische in 5195) [ClassicSimilarity], result of:
            0.08455382 = score(doc=5195,freq=1.0), product of:
              0.19506317 = queryWeight, product of:
                1.5811883 = boost
                6.9355025 = idf(docFreq=112, maxDocs=42740)
                0.017787453 = queryNorm
              0.4334689 = fieldWeight in 5195, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.9355025 = idf(docFreq=112, maxDocs=42740)
                0.0625 = fieldNorm(doc=5195)
          0.12973589 = weight(abstract_txt:statistischen in 5195) [ClassicSimilarity], result of:
            0.12973589 = score(doc=5195,freq=1.0), product of:
              0.25949353 = queryWeight, product of:
                1.8237245 = boost
                7.999329 = idf(docFreq=38, maxDocs=42740)
                0.017787453 = queryNorm
              0.49995807 = fieldWeight in 5195, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.999329 = idf(docFreq=38, maxDocs=42740)
                0.0625 = fieldNorm(doc=5195)
          0.22308224 = weight(abstract_txt:indexterme in 5195) [ClassicSimilarity], result of:
            0.22308224 = score(doc=5195,freq=1.0), product of:
              0.37244585 = queryWeight, product of:
                2.1848798 = boost
                9.583449 = idf(docFreq=7, maxDocs=42740)
                0.017787453 = queryNorm
              0.5989656 = fieldWeight in 5195, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.583449 = idf(docFreq=7, maxDocs=42740)
                0.0625 = fieldNorm(doc=5195)
          0.0815516 = weight(abstract_txt:werden in 5195) [ClassicSimilarity], result of:
            0.0815516 = score(doc=5195,freq=6.0), product of:
              0.15113491 = queryWeight, product of:
                2.4106767 = boost
                3.524618 = idf(docFreq=3422, maxDocs=42740)
                0.017787453 = queryNorm
              0.5395947 = fieldWeight in 5195, product of:
                2.4494898 = tf(freq=6.0), with freq of:
                  6.0 = termFreq=6.0
                3.524618 = idf(docFreq=3422, maxDocs=42740)
                0.0625 = fieldNorm(doc=5195)
        0.24 = coord(6/25)
    
  3. Leonhardt, H.A.: Systematik "Ästhetische Kulturwissenschaft" an der Universitätsbibliothek Hildesheim : ein Innovationsbericht (2018) 0.13
    0.13210776 = sum of:
      0.13210776 = product of:
        0.6605388 = sum of:
          0.033367958 = weight(abstract_txt:eine in 491) [ClassicSimilarity], result of:
            0.033367958 = score(doc=491,freq=1.0), product of:
              0.07568038 = queryWeight, product of:
                1.2062386 = boost
                3.5272505 = idf(docFreq=3413, maxDocs=42740)
                0.017787453 = queryNorm
              0.44090632 = fieldWeight in 491, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                3.5272505 = idf(docFreq=3413, maxDocs=42740)
                0.125 = fieldNorm(doc=491)
          0.15151621 = weight(abstract_txt:bücher in 491) [ClassicSimilarity], result of:
            0.15151621 = score(doc=491,freq=1.0), product of:
              0.18128945 = queryWeight, product of:
                1.5243413 = boost
                6.6861567 = idf(docFreq=144, maxDocs=42740)
                0.017787453 = queryNorm
              0.8357696 = fieldWeight in 491, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.6861567 = idf(docFreq=144, maxDocs=42740)
                0.125 = fieldNorm(doc=491)
          0.25232694 = weight(abstract_txt:indexieren in 491) [ClassicSimilarity], result of:
            0.25232694 = score(doc=491,freq=1.0), product of:
              0.25470778 = queryWeight, product of:
                1.806829 = boost
                7.925221 = idf(docFreq=41, maxDocs=42740)
                0.017787453 = queryNorm
              0.9906526 = fieldWeight in 491, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.925221 = idf(docFreq=41, maxDocs=42740)
                0.125 = fieldNorm(doc=491)
          0.12916006 = weight(abstract_txt:teil in 491) [ClassicSimilarity], result of:
            0.12916006 = score(doc=491,freq=1.0), product of:
              0.1865731 = queryWeight, product of:
                1.8939395 = boost
                5.538207 = idf(docFreq=456, maxDocs=42740)
                0.017787453 = queryNorm
              0.6922759 = fieldWeight in 491, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.538207 = idf(docFreq=456, maxDocs=42740)
                0.125 = fieldNorm(doc=491)
          0.09416767 = weight(abstract_txt:werden in 491) [ClassicSimilarity], result of:
            0.09416767 = score(doc=491,freq=2.0), product of:
              0.15113491 = queryWeight, product of:
                2.4106767 = boost
                3.524618 = idf(docFreq=3422, maxDocs=42740)
                0.017787453 = queryNorm
              0.6230703 = fieldWeight in 491, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.524618 = idf(docFreq=3422, maxDocs=42740)
                0.125 = fieldNorm(doc=491)
        0.2 = coord(5/25)
    
  4. Larroche-Boutet, V.; Pöhl, K.: ¬Das Nominalsyntagna : über die Nutzbarmachung eines logico-semantischen Konzeptes für dokumentarische Fragestellungen (1993) 0.12
    0.118032746 = sum of:
      0.118032746 = product of:
        0.5901637 = sum of:
          0.095433064 = weight(abstract_txt:extrahiert in 198) [ClassicSimilarity], result of:
            0.095433064 = score(doc=198,freq=1.0), product of:
              0.16783236 = queryWeight, product of:
                1.0370957 = boost
                9.097941 = idf(docFreq=12, maxDocs=42740)
                0.017787453 = queryNorm
              0.56862134 = fieldWeight in 198, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                9.097941 = idf(docFreq=12, maxDocs=42740)
                0.0625 = fieldNorm(doc=198)
          0.02359471 = weight(abstract_txt:eine in 198) [ClassicSimilarity], result of:
            0.02359471 = score(doc=198,freq=2.0), product of:
              0.07568038 = queryWeight, product of:
                1.2062386 = boost
                3.5272505 = idf(docFreq=3413, maxDocs=42740)
                0.017787453 = queryNorm
              0.31176785 = fieldWeight in 198, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                3.5272505 = idf(docFreq=3413, maxDocs=42740)
                0.0625 = fieldNorm(doc=198)
          0.11957716 = weight(abstract_txt:automatische in 198) [ClassicSimilarity], result of:
            0.11957716 = score(doc=198,freq=2.0), product of:
              0.19506317 = queryWeight, product of:
                1.5811883 = boost
                6.9355025 = idf(docFreq=112, maxDocs=42740)
                0.017787453 = queryNorm
              0.6130176 = fieldWeight in 198, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                6.9355025 = idf(docFreq=112, maxDocs=42740)
                0.0625 = fieldNorm(doc=198)
          0.27711266 = weight(abstract_txt:indexierungsverfahren in 198) [ClassicSimilarity], result of:
            0.27711266 = score(doc=198,freq=2.0), product of:
              0.341597 = queryWeight, product of:
                2.09244 = boost
                9.177984 = idf(docFreq=11, maxDocs=42740)
                0.017787453 = queryNorm
              0.81122684 = fieldWeight in 198, product of:
                1.4142135 = tf(freq=2.0), with freq of:
                  2.0 = termFreq=2.0
                9.177984 = idf(docFreq=11, maxDocs=42740)
                0.0625 = fieldNorm(doc=198)
          0.07444608 = weight(abstract_txt:werden in 198) [ClassicSimilarity], result of:
            0.07444608 = score(doc=198,freq=5.0), product of:
              0.15113491 = queryWeight, product of:
                2.4106767 = boost
                3.524618 = idf(docFreq=3422, maxDocs=42740)
                0.017787453 = queryNorm
              0.49258032 = fieldWeight in 198, product of:
                2.236068 = tf(freq=5.0), with freq of:
                  5.0 = termFreq=5.0
                3.524618 = idf(docFreq=3422, maxDocs=42740)
                0.0625 = fieldNorm(doc=198)
        0.2 = coord(5/25)
    
  5. Peters, G.; Gaese, V.: ¬Das DocCat-System in der Textdokumentation von G+J (2003) 0.11
    0.11213717 = sum of:
      0.11213717 = product of:
        0.4004899 = sum of:
          0.035392065 = weight(abstract_txt:eine in 2508) [ClassicSimilarity], result of:
            0.035392065 = score(doc=2508,freq=8.0), product of:
              0.07568038 = queryWeight, product of:
                1.2062386 = boost
                3.5272505 = idf(docFreq=3413, maxDocs=42740)
                0.017787453 = queryNorm
              0.46765178 = fieldWeight in 2508, product of:
                2.828427 = tf(freq=8.0), with freq of:
                  8.0 = termFreq=8.0
                3.5272505 = idf(docFreq=3413, maxDocs=42740)
                0.046875 = fieldNorm(doc=2508)
          0.051219236 = weight(abstract_txt:dokumenten in 2508) [ClassicSimilarity], result of:
            0.051219236 = score(doc=2508,freq=1.0), product of:
              0.16917428 = queryWeight, product of:
                1.4725266 = boost
                6.458884 = idf(docFreq=181, maxDocs=42740)
                0.017787453 = queryNorm
              0.30276018 = fieldWeight in 2508, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.458884 = idf(docFreq=181, maxDocs=42740)
                0.046875 = fieldNorm(doc=2508)
          0.06341536 = weight(abstract_txt:automatische in 2508) [ClassicSimilarity], result of:
            0.06341536 = score(doc=2508,freq=1.0), product of:
              0.19506317 = queryWeight, product of:
                1.5811883 = boost
                6.9355025 = idf(docFreq=112, maxDocs=42740)
                0.017787453 = queryNorm
              0.32510167 = fieldWeight in 2508, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.9355025 = idf(docFreq=112, maxDocs=42740)
                0.046875 = fieldNorm(doc=2508)
          0.06415632 = weight(abstract_txt:methode in 2508) [ClassicSimilarity], result of:
            0.06415632 = score(doc=2508,freq=1.0), product of:
              0.19657966 = queryWeight, product of:
                1.5873228 = boost
                6.96241 = idf(docFreq=109, maxDocs=42740)
                0.017787453 = queryNorm
              0.32636297 = fieldWeight in 2508, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                6.96241 = idf(docFreq=109, maxDocs=42740)
                0.046875 = fieldNorm(doc=2508)
          0.094622605 = weight(abstract_txt:indexieren in 2508) [ClassicSimilarity], result of:
            0.094622605 = score(doc=2508,freq=1.0), product of:
              0.25470778 = queryWeight, product of:
                1.806829 = boost
                7.925221 = idf(docFreq=41, maxDocs=42740)
                0.017787453 = queryNorm
              0.37149474 = fieldWeight in 2508, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                7.925221 = idf(docFreq=41, maxDocs=42740)
                0.046875 = fieldNorm(doc=2508)
          0.04843502 = weight(abstract_txt:teil in 2508) [ClassicSimilarity], result of:
            0.04843502 = score(doc=2508,freq=1.0), product of:
              0.1865731 = queryWeight, product of:
                1.8939395 = boost
                5.538207 = idf(docFreq=456, maxDocs=42740)
                0.017787453 = queryNorm
              0.25960344 = fieldWeight in 2508, product of:
                1.0 = tf(freq=1.0), with freq of:
                  1.0 = termFreq=1.0
                5.538207 = idf(docFreq=456, maxDocs=42740)
                0.046875 = fieldNorm(doc=2508)
          0.043249268 = weight(abstract_txt:werden in 2508) [ClassicSimilarity], result of:
            0.043249268 = score(doc=2508,freq=3.0), product of:
              0.15113491 = queryWeight, product of:
                2.4106767 = boost
                3.524618 = idf(docFreq=3422, maxDocs=42740)
                0.017787453 = queryNorm
              0.28616333 = fieldWeight in 2508, product of:
                1.7320508 = tf(freq=3.0), with freq of:
                  3.0 = termFreq=3.0
                3.524618 = idf(docFreq=3422, maxDocs=42740)
                0.046875 = fieldNorm(doc=2508)
        0.28 = coord(7/25)