Search (20 results, page 1 of 1)

Haag, M.: Automatic text summarization : Evaluation des Copernic Summarizer und mögliche Einsatzfelder in der Fachinformation der DaimlerCrysler AG (2002) 0.06

0.06347248 = product of:
  0.1586812 = sum of:
    0.017148608 = weight(_text_:23 in 649) [ClassicSimilarity], result of:
      0.017148608 = score(doc=649,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.23759183 = fieldWeight in 649, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.046875 = fieldNorm(doc=649)
    0.017148608 = weight(_text_:23 in 649) [ClassicSimilarity], result of:
      0.017148608 = score(doc=649,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.23759183 = fieldWeight in 649, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.046875 = fieldNorm(doc=649)
    0.029713312 = weight(_text_:software in 649) [ClassicSimilarity], result of:
      0.029713312 = score(doc=649,freq=4.0), product of:
        0.07989157 = queryWeight, product of:
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02013827 = queryNorm
        0.3719205 = fieldWeight in 649, product of:
          2.0 = tf(freq=4.0), with freq of:
            4.0 = termFreq=4.0
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.046875 = fieldNorm(doc=649)
    0.0065578544 = weight(_text_:und in 649) [ClassicSimilarity], result of:
      0.0065578544 = score(doc=649,freq=2.0), product of:
        0.044633795 = queryWeight, product of:
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.02013827 = queryNorm
        0.14692576 = fieldWeight in 649, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.046875 = fieldNorm(doc=649)
    0.017148608 = weight(_text_:23 in 649) [ClassicSimilarity], result of:
      0.017148608 = score(doc=649,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.23759183 = fieldWeight in 649, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.046875 = fieldNorm(doc=649)
    0.029713312 = weight(_text_:software in 649) [ClassicSimilarity], result of:
      0.029713312 = score(doc=649,freq=4.0), product of:
        0.07989157 = queryWeight, product of:
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02013827 = queryNorm
        0.3719205 = fieldWeight in 649, product of:
          2.0 = tf(freq=4.0), with freq of:
            4.0 = termFreq=4.0
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.046875 = fieldNorm(doc=649)
    0.011537581 = weight(_text_:der in 649) [ClassicSimilarity], result of:
      0.011537581 = score(doc=649,freq=6.0), product of:
        0.044984195 = queryWeight, product of:
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.02013827 = queryNorm
        0.25648075 = fieldWeight in 649, product of:
          2.4494898 = tf(freq=6.0), with freq of:
            6.0 = termFreq=6.0
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.046875 = fieldNorm(doc=649)
    0.029713312 = weight(_text_:software in 649) [ClassicSimilarity], result of:
      0.029713312 = score(doc=649,freq=4.0), product of:
        0.07989157 = queryWeight, product of:
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02013827 = queryNorm
        0.3719205 = fieldWeight in 649, product of:
          2.0 = tf(freq=4.0), with freq of:
            4.0 = termFreq=4.0
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.046875 = fieldNorm(doc=649)
  0.4 = coord(8/20)

Abstract: An evaluation of the Copernic Summarizer, a software for automatically summarizing text in various data formats, is being presented. It shall be assessed if and how the Copernic Summarizer can reasonably be used in the DaimlerChrysler Information Division in order to enhance the quality of its information services. First, an introduction into Automatic Text Summarization is given and the Copernic Summarizer is being presented. Various methods for evaluating Automatic Text Summarization systems and software ergonomics are presented. Two evaluation forms are developed with which the employees of the Information Division shall evaluate the quality and relevance of the extracted keywords and summaries as well as the software's usability. The quality and relevance assessment is done by comparing the original text to the summaries. Finally, a recommendation is given concerning the use of the Copernic Summarizer.
Date: 27. 2.1999 16:23:24
Footnote: Diplomarbeit an der HBI Stuttgart. - Vgl. auch: nfd 53(2002) H.4, S.243-244

Meyer, R.: Allein, es wär' so schön gewesen : Der Copernic Summarzier kann Internettexte leider nicht befriedigend und sinnvoll zusammenfassen (2002) 0.02
```
0.024851477 = product of:
  0.09940591 = sum of:
    0.02451223 = weight(_text_:software in 648) [ClassicSimilarity], result of:
      0.02451223 = score(doc=648,freq=8.0), product of:
        0.07989157 = queryWeight, product of:
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02013827 = queryNorm
        0.30681872 = fieldWeight in 648, product of:
          2.828427 = tf(freq=8.0), with freq of:
            8.0 = termFreq=8.0
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02734375 = fieldNorm(doc=648)
    0.010819908 = weight(_text_:und in 648) [ClassicSimilarity], result of:
      0.010819908 = score(doc=648,freq=16.0), product of:
        0.044633795 = queryWeight, product of:
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.02013827 = queryNorm
        0.24241515 = fieldWeight in 648, product of:
          4.0 = tf(freq=16.0), with freq of:
            16.0 = termFreq=16.0
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.02734375 = fieldNorm(doc=648)
    0.02451223 = weight(_text_:software in 648) [ClassicSimilarity], result of:
      0.02451223 = score(doc=648,freq=8.0), product of:
        0.07989157 = queryWeight, product of:
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02013827 = queryNorm
        0.30681872 = fieldWeight in 648, product of:
          2.828427 = tf(freq=8.0), with freq of:
            8.0 = termFreq=8.0
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02734375 = fieldNorm(doc=648)
    0.015049307 = weight(_text_:der in 648) [ClassicSimilarity], result of:
      0.015049307 = score(doc=648,freq=30.0), product of:
        0.044984195 = queryWeight, product of:
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.02013827 = queryNorm
        0.33454654 = fieldWeight in 648, product of:
          5.477226 = tf(freq=30.0), with freq of:
            30.0 = termFreq=30.0
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.02734375 = fieldNorm(doc=648)
    0.02451223 = weight(_text_:software in 648) [ClassicSimilarity], result of:
      0.02451223 = score(doc=648,freq=8.0), product of:
        0.07989157 = queryWeight, product of:
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02013827 = queryNorm
        0.30681872 = fieldWeight in 648, product of:
          2.828427 = tf(freq=8.0), with freq of:
            8.0 = termFreq=8.0
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02734375 = fieldNorm(doc=648)
  0.25 = coord(5/20)
```
Abstract

Das Netz hat die Jagd nach textlichen Inhalten erheblich erleichtert. Es ist so ein-fach, irgendeinen Beitrag über ein bestimmtes Thema zu finden, daß man eher über Fülle als über Mangel klagt. Suchmaschinen und Kataloge helfen beim Sichten, indem sie eine Vorauswahl von Links treffen. Das Programm "Copernic Summarizer" geht einen anderen Weg: Es erstellt Exzerpte beliebiger Texte und will damit die Lesezeit verkürzen. Decken wir über die lästige Zwangsregistrierung (unter Pflichtangabe einer Mailadresse) das Mäntelchen des Schweigens. Was folgt, geht rasch, nicht nur die ersten Schritte sind schnell vollzogen. Die Software läßt sich in verschiedenen Umgebungen einsetzen. Unterstützt werden Microsoft Office, einige Mailprogramme sowie der Acrobat Reader für PDF-Dateien. Besonders eignet sich das Verfahren freilich für Internetseiten. Der "Summarizer" nistet sich im Browser als Symbol ein. Und mit einem Klick faßt er einen Online Text in einem Extrafenster zusammen. Es handelt sich dabei nicht im eigentlichen Sinne um eine Zusammenfassung mit eigenen Worten, die in Kürze den Gesamtgehalt wiedergibt. Das Ergebnis ist schlichtes Kürzen, das sich noch dazu ziemlich brutal vollzieht, da grundsätzlich vollständige Sätze gestrichen werden. Die Software erfaßt den Text, versucht Schlüsselwörter zu ermitteln und entscheidet danach, welche Sätze wichtig sind und welche nicht. Das Verfahren mag den Entwicklungsaufwand verringert haben, dem Anwender hingegen bereitet es Probleme. Oftmals beziehen sich Sätze auf frühere Aussagen, etwa in Formulierungen wie "Diese Methode wird . . ." oder "Ein Jahr später . . ." In der Zusammenfassung fehlt entweder der Kontext dazu oder man kann nicht darauf vertrauen, daß der Bezug sich tatsächlich im voranstehenden Satz findet. Die Liste der Schlüsselwörter, die links eingeblendet wird, wirkt nicht immer glücklich. Teilweise finden sich unauffällige Begriffe wie "Anlaß" oder "zudem". Wenigstens lassen sich einzelne Begriffe entfernen, um das Ergebnis zu verfeinern. Hilfreich ist das mögliche Markieren der Schlüsselbegriffe im Text. Unverständlich bleibt hingegen, weshalb man nicht selbst relevante Wörter festlegen darf, die als Basis für die Zusammenfassung dienen. Das Kürzen des Textes ist in mehreren Stufen möglich, von fünf bis fünfzig Prozent. Fünf Prozent sind unbrauchbar; ein guter Kompromiß sind fünfundzwanzig. Allerdings nimmt es die Software nicht genau mit den eigenen Vorgaben. Bei kürzeren Texten ist die Zusammenfassung von angeblich einem Viertel fast genauso lang wie das Original; noch bei zwei Seiten eng bedrucktem Text (8 Kilobyte) entspricht das Exzerpt einem Drittel des Originals. Für gewöhnlich sind Webseiten geschmückt mit einem Menü, mit Werbung, mit Hinweiskästen und allerlei mehr. Sehr zuverlässig erkennt die Software, was überhaupt Fließtext ist; alles andere wird ausgefiltert. Da bedauert man es zuweilen, daß der Summarizer nicht den kompletten Text listet, damit er in einer angenehmen Umgebung schwarz auf weiß gelesen oder gedruckt wird. Wahlweise zum manuellen Auslösen der Zusammenfassung wird der "LiveSummarizer" aktiviert. Er verdichtet Text zeitgleich mit dem Aufrufen einer Seite, nimmt dafür aber ein Drittel der Bildschirmfläche ein - ein zu hoher Preis. Insgesamt fragen wir uns, wie man das Programm sinnvoll nutzen soll. Beim Verdichten von Nachrichten ist unsicher, ob Summarizer nicht wichtige Details unterschlägt. Bei langen Texten sorgen Fragen zum Kontext für Verwirrung. Sucht man nach der Antwort auf eine Detailfrage, hilft die Suchfunktion des Browsers oft schneller. Eine Zusammenfassung hätte auch dem Preis gutgetan: 100 Euro verlangt der deutsche Verleger Softline. Das scheint deutlich zu hoch gegriffen. Zumal das Zusammenfassen der einzige Zweck des Summarizers ist. Das Verwalten von Bookmarks und das Archivieren von Texten wären sinnvolle Ergänzungen gewesen.

Kuhlen, R.: In Richtung Summarizing für Diskurse in K3 (2006) 0.02

0.023626339 = product of:
  0.094505355 = sum of:
    0.02000671 = weight(_text_:23 in 6067) [ClassicSimilarity], result of:
      0.02000671 = score(doc=6067,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.27719048 = fieldWeight in 6067, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.0546875 = fieldNorm(doc=6067)
    0.02000671 = weight(_text_:23 in 6067) [ClassicSimilarity], result of:
      0.02000671 = score(doc=6067,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.27719048 = fieldWeight in 6067, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.0546875 = fieldNorm(doc=6067)
    0.017107777 = weight(_text_:und in 6067) [ClassicSimilarity], result of:
      0.017107777 = score(doc=6067,freq=10.0), product of:
        0.044633795 = queryWeight, product of:
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.02013827 = queryNorm
        0.38329202 = fieldWeight in 6067, product of:
          3.1622777 = tf(freq=10.0), with freq of:
            10.0 = termFreq=10.0
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.0546875 = fieldNorm(doc=6067)
    0.02000671 = weight(_text_:23 in 6067) [ClassicSimilarity], result of:
      0.02000671 = score(doc=6067,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.27719048 = fieldWeight in 6067, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.0546875 = fieldNorm(doc=6067)
    0.017377444 = weight(_text_:der in 6067) [ClassicSimilarity], result of:
      0.017377444 = score(doc=6067,freq=10.0), product of:
        0.044984195 = queryWeight, product of:
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.02013827 = queryNorm
        0.38630107 = fieldWeight in 6067, product of:
          3.1622777 = tf(freq=10.0), with freq of:
            10.0 = termFreq=10.0
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.0546875 = fieldNorm(doc=6067)
  0.25 = coord(5/20)

Abstract: Der Bedarf nach Summarizing-Leistungen, in Situationen der Fachinformation, aber auch in kommunikativen Umgebungen (Diskursen) wird aufgezeigt. Summarizing wird dazu in den Kontext des bisherigen (auch automatischen) Abstracting/Extracting gestellt. Der aktuelle Forschungsstand, vor allem mit Blick auf Multi-Document-Summarizing, wird dargestellt. Summarizing ist eine wichtige Funktion in komplex und umfänglich werdenden Diskussionen in elektronischen Foren. Dies wird am Beispiel des e-Learning-Systems K3 aufgezeigt. Rudimentäre Summarizing-Funktionen von K3 und des zugeordneten K3VIS-Systems werden dargestellt. Der Rahmen für ein elaborierteres, Template-orientiertes Summarizing unter Verwendung der vielfältigen Auszeichnungsfunktionen von K3 (Rollen, Diskurstypen, Inhaltstypen etc.) wird aufgespannt.
Date: 13.10.2006 9:35:23
Source: Information und Sprache: Beiträge zu Informationswissenschaft, Computerlinguistik, Bibliothekswesen und verwandten Fächern. Festschrift für Harald H. Zimmermann. Herausgegeben von Ilse Harms, Heinz-Dirk Luckhardt und Hans W. Giessen

Craven, T.C.: Abstracts produced using computer assistance (2000) 0.01

0.013370992 = product of:
  0.08913994 = sum of:
    0.029713312 = weight(_text_:software in 4809) [ClassicSimilarity], result of:
      0.029713312 = score(doc=4809,freq=4.0), product of:
        0.07989157 = queryWeight, product of:
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02013827 = queryNorm
        0.3719205 = fieldWeight in 4809, product of:
          2.0 = tf(freq=4.0), with freq of:
            4.0 = termFreq=4.0
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.046875 = fieldNorm(doc=4809)
    0.029713312 = weight(_text_:software in 4809) [ClassicSimilarity], result of:
      0.029713312 = score(doc=4809,freq=4.0), product of:
        0.07989157 = queryWeight, product of:
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02013827 = queryNorm
        0.3719205 = fieldWeight in 4809, product of:
          2.0 = tf(freq=4.0), with freq of:
            4.0 = termFreq=4.0
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.046875 = fieldNorm(doc=4809)
    0.029713312 = weight(_text_:software in 4809) [ClassicSimilarity], result of:
      0.029713312 = score(doc=4809,freq=4.0), product of:
        0.07989157 = queryWeight, product of:
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02013827 = queryNorm
        0.3719205 = fieldWeight in 4809, product of:
          2.0 = tf(freq=4.0), with freq of:
            4.0 = termFreq=4.0
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.046875 = fieldNorm(doc=4809)
  0.15 = coord(3/20)

Abstract: Experimental subjects wrote abstracts using a simplified version of the TEXNET abstracting assistance software. In addition to the full text, subjects were presented with either keywords or phrases extracted automatically. The resulting abstracts, and the times taken, were recorded automatically; some additional information was gathered by oral questionnaire. Selected abstracts produced were evaluated on various criteria by independent raters. Results showed considerable variation among subjects, but 37% found the keywords or phrases 'quite' or 'very' useful in writing their abstracts. Statistical analysis failed to support several hypothesized relations: phrases were not viewed as significantly more helpful than keywords; and abstracting experience did not correlate with originality of wording, approximation of the author abstract, or greater conciseness. Requiring further study are some unanticipated strong correlations including the following: Windows experience and writing an abstract like the author's; experience reading abstracts and thinking one had written a good abstract; gender and abstract length; gender and use of words and phrases from the original text. Results have also suggested possible modifications to the TEXNET software

Lee, J.-H.; Park, S.; Ahn, C.-M.; Kim, D.: Automatic generic document summarization based on non-negative matrix factorization (2009) 0.01

0.009003021 = product of:
  0.060020134 = sum of:
    0.02000671 = weight(_text_:23 in 2448) [ClassicSimilarity], result of:
      0.02000671 = score(doc=2448,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.27719048 = fieldWeight in 2448, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.0546875 = fieldNorm(doc=2448)
    0.02000671 = weight(_text_:23 in 2448) [ClassicSimilarity], result of:
      0.02000671 = score(doc=2448,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.27719048 = fieldWeight in 2448, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.0546875 = fieldNorm(doc=2448)
    0.02000671 = weight(_text_:23 in 2448) [ClassicSimilarity], result of:
      0.02000671 = score(doc=2448,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.27719048 = fieldWeight in 2448, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.0546875 = fieldNorm(doc=2448)
  0.15 = coord(3/20)

Date: 23. 3.2013 13:24:19

Sparck Jones, K.: Automatic summarising : the state of the art (2007) 0.01

0.007716874 = product of:
  0.051445827 = sum of:
    0.017148608 = weight(_text_:23 in 932) [ClassicSimilarity], result of:
      0.017148608 = score(doc=932,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.23759183 = fieldWeight in 932, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.046875 = fieldNorm(doc=932)
    0.017148608 = weight(_text_:23 in 932) [ClassicSimilarity], result of:
      0.017148608 = score(doc=932,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.23759183 = fieldWeight in 932, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.046875 = fieldNorm(doc=932)
    0.017148608 = weight(_text_:23 in 932) [ClassicSimilarity], result of:
      0.017148608 = score(doc=932,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.23759183 = fieldWeight in 932, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.046875 = fieldNorm(doc=932)
  0.15 = coord(3/20)

Date: 26.12.2007 14:40:23

Díaz, A.; Gervás, P.: User-model based personalized summarization (2007) 0.01

0.007716874 = product of:
  0.051445827 = sum of:
    0.017148608 = weight(_text_:23 in 952) [ClassicSimilarity], result of:
      0.017148608 = score(doc=952,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.23759183 = fieldWeight in 952, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.046875 = fieldNorm(doc=952)
    0.017148608 = weight(_text_:23 in 952) [ClassicSimilarity], result of:
      0.017148608 = score(doc=952,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.23759183 = fieldWeight in 952, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.046875 = fieldNorm(doc=952)
    0.017148608 = weight(_text_:23 in 952) [ClassicSimilarity], result of:
      0.017148608 = score(doc=952,freq=2.0), product of:
        0.07217676 = queryWeight, product of:
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.02013827 = queryNorm
        0.23759183 = fieldWeight in 952, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.5840597 = idf(docFreq=3336, maxDocs=44218)
          0.046875 = fieldNorm(doc=952)
  0.15 = coord(3/20)

Date: 26.12.2007 16:27:23

Pinto, M.: Engineering the production of meta-information : the abstracting concern (2003) 0.01

0.0072211013 = product of:
  0.07221101 = sum of:
    0.027255533 = product of:
      0.054511067 = sum of:
        0.054511067 = weight(_text_:29 in 4667) [ClassicSimilarity], result of:
          0.054511067 = score(doc=4667,freq=4.0), product of:
            0.070840135 = queryWeight, product of:
              3.5176873 = idf(docFreq=3565, maxDocs=44218)
              0.02013827 = queryNorm
            0.7694941 = fieldWeight in 4667, product of:
              2.0 = tf(freq=4.0), with freq of:
                4.0 = termFreq=4.0
              3.5176873 = idf(docFreq=3565, maxDocs=44218)
              0.109375 = fieldNorm(doc=4667)
      0.5 = coord(1/2)
    0.044955477 = product of:
      0.089910954 = sum of:
        0.089910954 = weight(_text_:engineering in 4667) [ClassicSimilarity], result of:
          0.089910954 = score(doc=4667,freq=2.0), product of:
            0.10819342 = queryWeight, product of:
              5.372528 = idf(docFreq=557, maxDocs=44218)
              0.02013827 = queryNorm
            0.83102053 = fieldWeight in 4667, product of:
              1.4142135 = tf(freq=2.0), with freq of:
                2.0 = termFreq=2.0
              5.372528 = idf(docFreq=557, maxDocs=44218)
              0.109375 = fieldNorm(doc=4667)
      0.5 = coord(1/2)
  0.1 = coord(2/20)

Date: 27.11.2005 18:29:55
Source: Journal of information science. 29(2003) no.5, S.405-418

Craven, T.C.: Presentation of repeated phrases in a computer-assisted abstracting tool kit (2001) 0.01

0.0070481906 = product of:
  0.14096381 = sum of:
    0.14096381 = weight(_text_:230 in 3667) [ClassicSimilarity], result of:
      0.14096381 = score(doc=3667,freq=2.0), product of:
        0.13547163 = queryWeight, product of:
          6.727074 = idf(docFreq=143, maxDocs=44218)
          0.02013827 = queryNorm
        1.0405412 = fieldWeight in 3667, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          6.727074 = idf(docFreq=143, maxDocs=44218)
          0.109375 = fieldNorm(doc=3667)
  0.05 = coord(1/20)

Source: Information processing and management. 37(2001) no.2, S.221-230

Dunlavy, D.M.; O'Leary, D.P.; Conroy, J.M.; Schlesinger, J.D.: QCS: A system for querying, clustering and summarizing documents (2007) 0.01
```
0.0063031456 = product of:
  0.04202097 = sum of:
    0.014006989 = weight(_text_:software in 947) [ClassicSimilarity], result of:
      0.014006989 = score(doc=947,freq=2.0), product of:
        0.07989157 = queryWeight, product of:
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02013827 = queryNorm
        0.17532499 = fieldWeight in 947, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.03125 = fieldNorm(doc=947)
    0.014006989 = weight(_text_:software in 947) [ClassicSimilarity], result of:
      0.014006989 = score(doc=947,freq=2.0), product of:
        0.07989157 = queryWeight, product of:
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02013827 = queryNorm
        0.17532499 = fieldWeight in 947, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.03125 = fieldNorm(doc=947)
    0.014006989 = weight(_text_:software in 947) [ClassicSimilarity], result of:
      0.014006989 = score(doc=947,freq=2.0), product of:
        0.07989157 = queryWeight, product of:
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.02013827 = queryNorm
        0.17532499 = fieldWeight in 947, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          3.9671519 = idf(docFreq=2274, maxDocs=44218)
          0.03125 = fieldNorm(doc=947)
  0.15 = coord(3/20)
```
Abstract

Information retrieval systems consist of many complicated components. Research and development of such systems is often hampered by the difficulty in evaluating how each particular component would behave across multiple systems. We present a novel integrated information retrieval system-the Query, Cluster, Summarize (QCS) system-which is portable, modular, and permits experimentation with different instantiations of each of the constituent text analysis components. Most importantly, the combination of the three types of methods in the QCS design improves retrievals by providing users more focused information organized by topic. We demonstrate the improved performance by a series of experiments using standard test sets from the Document Understanding Conferences (DUC) as measured by the best known automatic metric for summarization system evaluation, ROUGE. Although the DUC data and evaluations were originally designed to test multidocument summarization, we developed a framework to extend it to the task of evaluation for each of the three components: query, clustering, and summarization. Under this framework, we then demonstrate that the QCS system (end-to-end) achieves performance as good as or better than the best summarization engines. Given a query, QCS retrieves relevant documents, separates the retrieved documents into topic clusters, and creates a single summary for each cluster. In the current implementation, Latent Semantic Indexing is used for retrieval, generalized spherical k-means is used for the document clustering, and a method coupling sentence "trimming" and a hidden Markov model, followed by a pivoted QR decomposition, is used to create a single extract summary for each cluster. The user interface is designed to provide access to detailed information in a compact and useful format. Our system demonstrates the feasibility of assembling an effective IR system from existing software libraries, the usefulness of the modularity of the design, and the value of this particular combination of modules.
Endres-Niggemeyer, B.; Jauris-Heipke, S.; Pinsky, S.M.; Ulbricht, U.: Wissen gewinnen durch Wissen : Ontologiebasierte Informationsextraktion (2006) 0.00
```
0.0028807095 = product of:
  0.028807094 = sum of:
    0.016394636 = weight(_text_:und in 6016) [ClassicSimilarity], result of:
      0.016394636 = score(doc=6016,freq=18.0), product of:
        0.044633795 = queryWeight, product of:
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.02013827 = queryNorm
        0.3673144 = fieldWeight in 6016, product of:
          4.2426405 = tf(freq=18.0), with freq of:
            18.0 = termFreq=18.0
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.0390625 = fieldNorm(doc=6016)
    0.012412459 = weight(_text_:der in 6016) [ClassicSimilarity], result of:
      0.012412459 = score(doc=6016,freq=10.0), product of:
        0.044984195 = queryWeight, product of:
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.02013827 = queryNorm
        0.27592933 = fieldWeight in 6016, product of:
          3.1622777 = tf(freq=10.0), with freq of:
            10.0 = termFreq=10.0
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.0390625 = fieldNorm(doc=6016)
  0.1 = coord(2/20)
```
Abstract

Die ontologiebasierte Informationsextraktion, über die hier berichtet wird, ist Teil eines Systems zum automatischen Zusammenfassen, das sich am Vorgehen kompetenter Menschen orientiert. Dahinter steht die Annahme, dass Menschen die Ergebnisse eines Systems leichter übernehmen können, wenn sie mit Verfahren erarbeitet worden sind, die sie selbst auch benutzen. Das erste Anwendungsgebiet ist Knochenmarktransplantation (KMT). Im Kern des Systems Summit-BMT (Summarize It in Bone Marrow Transplantation) steht eine Ontologie des Fachgebietes. Sie ist als MySQL-Datenbank realisiert und versorgt menschliche Benutzer und Systemkomponenten mit Wissen. Summit-BMT unterstützt die Frageformulierung mit einem empirisch fundierten Szenario-Interface. Die Retrievalergebnisse werden durch ein Textpassagenretrieval vorselektiert und dann kognitiv fundierten Agenten unterbreitet, die unter Einsatz ihrer Wissensbasis / Ontologie genauer prüfen, ob die Propositionen aus der Benutzerfrage getroffen werden. Die relevanten Textclips aus dem Duelldokument werden in das Szenarioformular eingetragen und mit einem Link zu ihrem Vorkommen im Original präsentiert. In diesem Artikel stehen die Ontologie und ihr Gebrauch zur wissensbasierten Informationsextraktion im Mittelpunkt. Die Ontologiedatenbank hält unterschiedliche Wissenstypen so bereit, dass sie leicht kombiniert werden können: Konzepte, Propositionen und ihre syntaktisch-semantischen Schemata, Unifikatoren, Paraphrasen und Definitionen von Frage-Szenarios. Auf sie stützen sich die Systemagenten, welche von Menschen adaptierte Zusammenfassungsstrategien ausführen. Mängel in anderen Verarbeitungsschritten führen zu Verlusten, aber die eigentliche Qualität der Ergebnisse steht und fällt mit der Qualität der Ontologie. Erste Tests der Extraktionsleistung fallen verblüffend positiv aus.

Source

Information - Wissenschaft und Praxis. 57(2006) H.6/7, S.301-308
Endres-Niggemeyer, B.; Ziegert, C.: SummIt-BMT : (Summarize It in BMT) in Diagnose und Therapie, Abschlussbericht (2002) 0.00
```
0.0026630415 = product of:
  0.026630415 = sum of:
    0.010929758 = weight(_text_:und in 4497) [ClassicSimilarity], result of:
      0.010929758 = score(doc=4497,freq=8.0), product of:
        0.044633795 = queryWeight, product of:
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.02013827 = queryNorm
        0.24487628 = fieldWeight in 4497, product of:
          2.828427 = tf(freq=8.0), with freq of:
            8.0 = termFreq=8.0
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.0390625 = fieldNorm(doc=4497)
    0.015700657 = weight(_text_:der in 4497) [ClassicSimilarity], result of:
      0.015700657 = score(doc=4497,freq=16.0), product of:
        0.044984195 = queryWeight, product of:
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.02013827 = queryNorm
        0.34902605 = fieldWeight in 4497, product of:
          4.0 = tf(freq=16.0), with freq of:
            16.0 = termFreq=16.0
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.0390625 = fieldNorm(doc=4497)
  0.1 = coord(2/20)
```
Abstract

SummIt-BMT (Summarize It in Bone Marrow Transplantation) - das Zielsystem des Projektes - soll Ärzten in der Knochenmarktransplantation durch kognitiv fundiertes Zusammenfassen (Endres-Niggemeyer, 1998) aus dem WWW eine schnelle Informationsaufnahme ermöglichen. Im bmbffinanzierten Teilprojekt, über das hier zu berichten ist, liegt der Schwerpunkt auf den klinischen Fragestellungen. SummIt-BMT hat als zentrale Komponente eine KMT-Ontologie. Den Systemablauf veranschaulicht Abb. 1: Benutzer geben ihren Informationsbedarf in ein strukturiertes Szenario ein. Sie ziehen dazu Begriffe aus der Ontologie heran. Aus dem Szenario werden Fragen an Suchmaschinen abgeleitet. Die Summit-BMT-Metasuchmaschine stößt Google an und sucht in Medline, der zentralen Literaturdatenbank der Medizin. Das Suchergebnis wird aufbereitet. Dabei werden Links zu Volltexten verfolgt und die Volltexte besorgt. Die beschafften Dokumente werden mit einem Schlüsselwortretrieval auf Passagen untersucht, in denen sich Suchkonzepte aus der Frage / Ontologie häufen. Diese Passagen werden zum Zusammenfassen vorgeschlagen. In ihnen werden die Aussagen syntaktisch analysiert. Die Systemagenten untersuchen sie. Lassen Aussagen sich mit einer semantischen Relation an die Frage anbinden, tragen also zur deren Beantwortung bei, werden sie in die Zusammenfassung aufgenommen, es sei denn, andere Agenten machen Hinderungsgründe geltend, z.B. Redundanz. Das Ergebnis der Zusammenfassung wird in das Frage/Antwort-Szenario integriert. Präsentiert werden Exzerpte aus den Quelldokumenten. Mit einem Link vermitteln sie einen sofortigen Rückgriff auf die Quelle. SummIt-BMT ist zum nächsten Durchgang von Informationssuche und Zusammenfassung bereit, sobald der Benutzer dies wünscht.

Haag, M.: Automatic text summarization (2002) 0.00

0.002643816 = product of:
  0.026438158 = sum of:
    0.013115709 = weight(_text_:und in 5662) [ClassicSimilarity], result of:
      0.013115709 = score(doc=5662,freq=2.0), product of:
        0.044633795 = queryWeight, product of:
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.02013827 = queryNorm
        0.29385152 = fieldWeight in 5662, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.09375 = fieldNorm(doc=5662)
    0.013322448 = weight(_text_:der in 5662) [ClassicSimilarity], result of:
      0.013322448 = score(doc=5662,freq=2.0), product of:
        0.044984195 = queryWeight, product of:
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.02013827 = queryNorm
        0.29615843 = fieldWeight in 5662, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.09375 = fieldNorm(doc=5662)
  0.1 = coord(2/20)

Footnote: Basiert auf einer Diplomabarbeit des Verfassers an der HBI Stuttgart, die im Shaker-Verlag erschienen ist
Source: Information - Wissenschaft und Praxis. 53(2002) H.4, 243-244

Kuhlen, R.: Informationsaufbereitung III : Referieren (Abstracts - Abstracting - Grundlagen) (2004) 0.00
```
0.0020493101 = product of:
  0.020493101 = sum of:
    0.008743806 = weight(_text_:und in 2917) [ClassicSimilarity], result of:
      0.008743806 = score(doc=2917,freq=8.0), product of:
        0.044633795 = queryWeight, product of:
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.02013827 = queryNorm
        0.19590102 = fieldWeight in 2917, product of:
          2.828427 = tf(freq=8.0), with freq of:
            8.0 = termFreq=8.0
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.03125 = fieldNorm(doc=2917)
    0.0117492955 = weight(_text_:der in 2917) [ClassicSimilarity], result of:
      0.0117492955 = score(doc=2917,freq=14.0), product of:
        0.044984195 = queryWeight, product of:
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.02013827 = queryNorm
        0.2611872 = fieldWeight in 2917, product of:
          3.7416575 = tf(freq=14.0), with freq of:
            14.0 = termFreq=14.0
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.03125 = fieldNorm(doc=2917)
  0.1 = coord(2/20)
```
Abstract

Was ein Abstract (im Folgenden synonym mit Referat oder Kurzreferat gebraucht) ist, legt das American National Standards Institute in einer Weise fest, die sicherlich von den meisten Fachleuten akzeptiert werden kann: "An abstract is defined as an abbreviated, accurate representation of the contents of a document"; fast genauso die deutsche Norm DIN 1426: "Das Kurzreferat gibt kurz und klar den Inhalt des Dokuments wieder." Abstracts gehören zum wissenschaftlichen Alltag. Weitgehend allen Publikationen, zumindest in den naturwissenschaftlichen, technischen, informationsbezogenen oder medizinischen Bereichen, gehen Abstracts voran, "prefe-rably prepared by its author(s) for publication with it". Es gibt wohl keinen Wissenschaftler, der nicht irgendwann einmal ein Abstract geschrieben hätte. Gehört das Erstellen von Abstracts dann überhaupt zur dokumentarischen bzw informationswissenschaftlichen Methodenlehre, wenn es jeder kann? Was macht den informationellen Mehrwert aus, der durch Expertenreferate gegenüber Laienreferaten erzeugt wird? Dies ist nicht so leicht zu beantworten, zumal geeignete Bewertungsverfahren fehlen, die Qualität von Abstracts vergleichend "objektiv" zu messen. Abstracts werden in erheblichem Umfang von Informationsspezialisten erstellt, oft unter der Annahme, dass Autoren selber dafür weniger geeignet sind. Vergegenwärtigen wir uns, was wir über Abstracts und Abstracting wissen. Ein besonders gelungenes Abstract ist zuweilen klarer als der Ursprungstext selber, darf aber nicht mehr Information als dieser enthalten: "Good abstracts are highly structured, concise, and coherent, and are the result of a thorough analysis of the content of the abstracted materials. Abstracts may be more readable than the basis documents, but because of size constraints they rarely equal and never surpass the information content of the basic document". Dies ist verständlich, denn ein "Abstract" ist zunächst nichts anderes als ein Ergebnis des Vorgangs einer Abstraktion. Ohne uns zu sehr in die philosophischen Hintergründe der Abstraktion zu verlieren, besteht diese doch "in der Vernachlässigung von bestimmten Vorstellungsbzw. Begriffsinhalten, von welchen zugunsten anderer Teilinhalte abgesehen, abstrahiert' wird. Sie ist stets verbunden mit einer Fixierung von (interessierenden) Merkmalen durch die aktive Aufmerksamkeit, die unter einem bestimmten pragmatischen Gesichtspunkt als wesentlich' für einen vorgestellten bzw für einen unter einen Begriff fallenden Gegenstand (oder eine Mehrheit von Gegenständen) betrachtet werden". Abstracts reduzieren weniger Begriffsinhalte, sondern Texte bezüglich ihres proportionalen Gehaltes. Borko/ Bernier haben dies sogar quantifiziert; sie schätzen den Reduktionsfaktor auf 1:10 bis 1:12

Source

Grundlagen der praktischen Information und Dokumentation. 5., völlig neu gefaßte Ausgabe. 2 Bde. Hrsg. von R. Kuhlen, Th. Seeger u. D. Strauch. Begründet von Klaus Laisiepen, Ernst Lutterbeck, Karl-Heinrich Meyer-Uhlenried. Bd.1: Handbuch zur Einführung in die Informationswissenschaft und -praxis

Hahn, U.: ¬Die Verdichtung textuellen Wissens zu Information : vom Wandel methodischer Paradigmen beim automatischen Abstracting (2004) 0.00

9.6146495E-4 = product of:
  0.019229298 = sum of:
    0.019229298 = weight(_text_:der in 4667) [ClassicSimilarity], result of:
      0.019229298 = score(doc=4667,freq=6.0), product of:
        0.044984195 = queryWeight, product of:
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.02013827 = queryNorm
        0.42746788 = fieldWeight in 4667, product of:
          2.4494898 = tf(freq=6.0), with freq of:
            6.0 = termFreq=6.0
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.078125 = fieldNorm(doc=4667)
  0.05 = coord(1/20)

Source: Wissen in Aktion: Der Primat der Pragmatik als Motto der Konstanzer Informationswissenschaft. Festschrift für Rainer Kuhlen. Hrsg. R. Hammwöhner, u.a

Yusuff, A.: Automatisches Indexing and Abstracting : Grundlagen und Beispiele (2002) 0.00

7.6508307E-4 = product of:
  0.015301661 = sum of:
    0.015301661 = weight(_text_:und in 1577) [ClassicSimilarity], result of:
      0.015301661 = score(doc=1577,freq=2.0), product of:
        0.044633795 = queryWeight, product of:
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.02013827 = queryNorm
        0.34282678 = fieldWeight in 1577, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          2.216367 = idf(docFreq=13101, maxDocs=44218)
          0.109375 = fieldNorm(doc=1577)
  0.05 = coord(1/20)

Deutsche Patentdatenbank mit maschinellen Abstract-Übersetzungen (2005) 0.00
```
4.440816E-4 = product of:
  0.008881632 = sum of:
    0.008881632 = weight(_text_:der in 3344) [ClassicSimilarity], result of:
      0.008881632 = score(doc=3344,freq=2.0), product of:
        0.044984195 = queryWeight, product of:
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.02013827 = queryNorm
        0.19743896 = fieldWeight in 3344, product of:
          1.4142135 = tf(freq=2.0), with freq of:
            2.0 = termFreq=2.0
          2.2337668 = idf(docFreq=12875, maxDocs=44218)
          0.0625 = fieldNorm(doc=3344)
  0.05 = coord(1/20)
```
Abstract

Dialog hat sein Angebot um eine deutsche Patentdatenbank des niederländischen Patentinformationsanbieters mit 3,1 Millionen Volltexten in deutscher Sprache ab 1967 ausgeweitet. Die Dokumente ab 1980 enthalten maschinell übersetzte Abstracts in englischer Sprache, die dem Deutschunkundigen einen ersten Überblick über die gefundenen Dokumente ermöglichen. Die Dokumente sind vier bis fünf Tage nach der Veröffentlichung durch das Deutsche Patentamt recherchierbar. Univentio bietet über Dialog auch WIPO/PCT im Volltext an.

Sweeney, S.; Crestani, F.; Losada, D.E.: 'Show me more' : incremental length summarisation using novelty detection (2008) 0.00

3.441531E-4 = product of:
  0.0068830615 = sum of:
    0.0068830615 = product of:
      0.013766123 = sum of:
        0.013766123 = weight(_text_:29 in 2054) [ClassicSimilarity], result of:
          0.013766123 = score(doc=2054,freq=2.0), product of:
            0.070840135 = queryWeight, product of:
              3.5176873 = idf(docFreq=3565, maxDocs=44218)
              0.02013827 = queryNorm
            0.19432661 = fieldWeight in 2054, product of:
              1.4142135 = tf(freq=2.0), with freq of:
                2.0 = termFreq=2.0
              3.5176873 = idf(docFreq=3565, maxDocs=44218)
              0.0390625 = fieldNorm(doc=2054)
      0.5 = coord(1/2)
  0.05 = coord(1/20)

Date: 29. 7.2008 19:35:12

Vanderwende, L.; Suzuki, H.; Brockett, J.M.; Nenkova, A.: Beyond SumBasic : task-focused summarization with sentence simplification and lexical expansion (2007) 0.00
```
2.7284576E-4 = product of:
  0.005456915 = sum of:
    0.005456915 = product of:
      0.016370745 = sum of:
        0.016370745 = weight(_text_:22 in 948) [ClassicSimilarity], result of:
          0.016370745 = score(doc=948,freq=2.0), product of:
            0.07052079 = queryWeight, product of:
              3.5018296 = idf(docFreq=3622, maxDocs=44218)
              0.02013827 = queryNorm
            0.23214069 = fieldWeight in 948, product of:
              1.4142135 = tf(freq=2.0), with freq of:
                2.0 = termFreq=2.0
              3.5018296 = idf(docFreq=3622, maxDocs=44218)
              0.046875 = fieldNorm(doc=948)
      0.33333334 = coord(1/3)
  0.05 = coord(1/20)
```
Abstract

In recent years, there has been increased interest in topic-focused multi-document summarization. In this task, automatic summaries are produced in response to a specific information request, or topic, stated by the user. The system we have designed to accomplish this task comprises four main components: a generic extractive summarization system, a topic-focusing component, sentence simplification, and lexical expansion of topic words. This paper details each of these components, together with experiments designed to quantify their individual contributions. We include an analysis of our results on two large datasets commonly used to evaluate task-focused summarization, the DUC2005 and DUC2006 datasets, using automatic metrics. Additionally, we include an analysis of our results on the DUC2006 task according to human evaluation metrics. In the human evaluation of system summaries compared to human summaries, i.e., the Pyramid method, our system ranked first out of 22 systems in terms of overall mean Pyramid score; and in the human evaluation of summary responsiveness to the topic, our system ranked third out of 35 systems.

Wu, Y.-f.B.; Li, Q.; Bot, R.S.; Chen, X.: Finding nuggets in documents : a machine learning approach (2006) 0.00

2.2737146E-4 = product of:
  0.0045474293 = sum of:
    0.0045474293 = product of:
      0.013642288 = sum of:
        0.013642288 = weight(_text_:22 in 5290) [ClassicSimilarity], result of:
          0.013642288 = score(doc=5290,freq=2.0), product of:
            0.07052079 = queryWeight, product of:
              3.5018296 = idf(docFreq=3622, maxDocs=44218)
              0.02013827 = queryNorm
            0.19345059 = fieldWeight in 5290, product of:
              1.4142135 = tf(freq=2.0), with freq of:
                2.0 = termFreq=2.0
              3.5018296 = idf(docFreq=3622, maxDocs=44218)
              0.0390625 = fieldNorm(doc=5290)
      0.33333334 = coord(1/3)
  0.05 = coord(1/20)

Date: 22. 7.2006 17:25:48

Search (20 results, page 1 of 1)

Authors

Languages

Types

Themes