Search (180 results, page 1 of 9)

Lackes, R.; Tillmanns, C.: Data Mining für die Unternehmenspraxis : Entscheidungshilfen und Fallstudien mit führenden Softwarelösungen (2006) 0.14
```
0.14013693 = product of:
  0.28027385 = sum of:
    0.28027385 = sum of:
      0.23908965 = weight(_text_:mining in 1383) [ClassicSimilarity], result of:
        0.23908965 = score(doc=1383,freq=10.0), product of:
          0.28585905 = queryWeight, product of:
            5.642448 = idf(docFreq=425, maxDocs=44218)
            0.05066224 = queryNorm
          0.83639 = fieldWeight in 1383, product of:
            3.1622777 = tf(freq=10.0), with freq of:
              10.0 = termFreq=10.0
            5.642448 = idf(docFreq=425, maxDocs=44218)
            0.046875 = fieldNorm(doc=1383)
      0.0411842 = weight(_text_:22 in 1383) [ClassicSimilarity], result of:
        0.0411842 = score(doc=1383,freq=2.0), product of:
          0.17741053 = queryWeight, product of:
            3.5018296 = idf(docFreq=3622, maxDocs=44218)
            0.05066224 = queryNorm
          0.23214069 = fieldWeight in 1383, product of:
            1.4142135 = tf(freq=2.0), with freq of:
              2.0 = termFreq=2.0
            3.5018296 = idf(docFreq=3622, maxDocs=44218)
            0.046875 = fieldNorm(doc=1383)
  0.5 = coord(1/2)
```
Abstract

Das Buch richtet sich an Praktiker in Unternehmen, die sich mit der Analyse von großen Datenbeständen beschäftigen. Nach einem kurzen Theorieteil werden vier Fallstudien aus dem Customer Relationship Management eines Versandhändlers bearbeitet. Dabei wurden acht führende Softwarelösungen verwendet: der Intelligent Miner von IBM, der Enterprise Miner von SAS, Clementine von SPSS, Knowledge Studio von Angoss, der Delta Miner von Bissantz, der Business Miner von Business Object und die Data Engine von MIT. Im Rahmen der Fallstudien werden die Stärken und Schwächen der einzelnen Lösungen deutlich, und die methodisch-korrekte Vorgehensweise beim Data Mining wird aufgezeigt. Beides liefert wertvolle Entscheidungshilfen für die Auswahl von Standardsoftware zum Data Mining und für die praktische Datenanalyse.

Content

Modelle, Methoden und Werkzeuge: Ziele und Aufbau der Untersuchung.- Grundlagen.- Planung und Entscheidung mit Data-Mining-Unterstützung.- Methoden.- Funktionalität und Handling der Softwarelösungen. Fallstudien: Ausgangssituation und Datenbestand im Versandhandel.- Kundensegmentierung.- Erklärung regionaler Marketingerfolge zur Neukundengewinnung.Prognose des Customer Lifetime Values.- Selektion von Kunden für eine Direktmarketingaktion.- Welche Softwarelösung für welche Entscheidung?- Fazit und Marktentwicklungen.

Date

22. 3.2008 14:46:06

Theme

Data Mining
Information visualization in data mining and knowledge discovery (2002) 0.10
```
0.10283139 = product of:
  0.20566279 = sum of:
    0.20566279 = sum of:
      0.19193472 = weight(_text_:mining in 1789) [ClassicSimilarity], result of:
        0.19193472 = score(doc=1789,freq=58.0), product of:
          0.28585905 = queryWeight, product of:
            5.642448 = idf(docFreq=425, maxDocs=44218)
            0.05066224 = queryNorm
          0.6714313 = fieldWeight in 1789, product of:
            7.615773 = tf(freq=58.0), with freq of:
              58.0 = termFreq=58.0
            5.642448 = idf(docFreq=425, maxDocs=44218)
            0.015625 = fieldNorm(doc=1789)
      0.013728068 = weight(_text_:22 in 1789) [ClassicSimilarity], result of:
        0.013728068 = score(doc=1789,freq=2.0), product of:
          0.17741053 = queryWeight, product of:
            3.5018296 = idf(docFreq=3622, maxDocs=44218)
            0.05066224 = queryNorm
          0.07738023 = fieldWeight in 1789, product of:
            1.4142135 = tf(freq=2.0), with freq of:
              2.0 = termFreq=2.0
            3.5018296 = idf(docFreq=3622, maxDocs=44218)
            0.015625 = fieldNorm(doc=1789)
  0.5 = coord(1/2)
```
Date

23. 3.2008 19:10:22

Footnote

Rez. in: JASIST 54(2003) no.9, S.905-906 (C.A. Badurek): "Visual approaches for knowledge discovery in very large databases are a prime research need for information scientists focused an extracting meaningful information from the ever growing stores of data from a variety of domains, including business, the geosciences, and satellite and medical imagery. This work presents a summary of research efforts in the fields of data mining, knowledge discovery, and data visualization with the goal of aiding the integration of research approaches and techniques from these major fields. The editors, leading computer scientists from academia and industry, present a collection of 32 papers from contributors who are incorporating visualization and data mining techniques through academic research as well application development in industry and government agencies. Information Visualization focuses upon techniques to enhance the natural abilities of humans to visually understand data, in particular, large-scale data sets. It is primarily concerned with developing interactive graphical representations to enable users to more intuitively make sense of multidimensional data as part of the data exploration process. It includes research from computer science, psychology, human-computer interaction, statistics, and information science. Knowledge Discovery in Databases (KDD) most often refers to the process of mining databases for previously unknown patterns and trends in data. Data mining refers to the particular computational methods or algorithms used in this process. The data mining research field is most related to computational advances in database theory, artificial intelligence and machine learning. This work compiles research summaries from these main research areas in order to provide "a reference work containing the collection of thoughts and ideas of noted researchers from the fields of data mining and data visualization" (p. 8). It addresses these areas in three main sections: the first an data visualization, the second an KDD and model visualization, and the last an using visualization in the knowledge discovery process. The seven chapters of Part One focus upon methodologies and successful techniques from the field of Data Visualization. Hoffman and Grinstein (Chapter 2) give a particularly good overview of the field of data visualization and its potential application to data mining. An introduction to the terminology of data visualization, relation to perceptual and cognitive science, and discussion of the major visualization display techniques are presented. Discussion and illustration explain the usefulness and proper context of such data visualization techniques as scatter plots, 2D and 3D isosurfaces, glyphs, parallel coordinates, and radial coordinate visualizations. Remaining chapters present the need for standardization of visualization methods, discussion of user requirements in the development of tools, and examples of using information visualization in addressing research problems.
In 13 chapters, Part Two provides an introduction to KDD, an overview of data mining techniques, and examples of the usefulness of data model visualizations. The importance of visualization throughout the KDD process is stressed in many of the chapters. In particular, the need for measures of visualization effectiveness, benchmarking for identifying best practices, and the use of standardized sample data sets is convincingly presented. Many of the important data mining approaches are discussed in this complementary context. Cluster and outlier detection, classification techniques, and rule discovery algorithms are presented as the basic techniques common to the KDD process. The potential effectiveness of using visualization in the data modeling process are illustrated in chapters focused an using visualization for helping users understand the KDD process, ask questions and form hypotheses about their data, and evaluate the accuracy and veracity of their results. The 11 chapters of Part Three provide an overview of the KDD process and successful approaches to integrating KDD, data mining, and visualization in complementary domains. Rhodes (Chapter 21) begins this section with an excellent overview of the relation between the KDD process and data mining techniques. He states that the "primary goals of data mining are to describe the existing data and to predict the behavior or characteristics of future data of the same type" (p. 281). These goals are met by data mining tasks such as classification, regression, clustering, summarization, dependency modeling, and change or deviation detection. Subsequent chapters demonstrate how visualization can aid users in the interactive process of knowledge discovery by graphically representing the results from these iterative tasks. Finally, examples of the usefulness of integrating visualization and data mining tools in the domain of business, imagery and text mining, and massive data sets are provided. This text concludes with a thorough and useful 17-page index and lengthy yet integrating 17-page summary of the academic and industrial backgrounds of the contributing authors. A 16-page set of color inserts provide a better representation of the visualizations discussed, and a URL provided suggests that readers may view all the book's figures in color on-line, although as of this submission date it only provides access to a summary of the book and its contents. The overall contribution of this work is its focus an bridging two distinct areas of research, making it a valuable addition to the Morgan Kaufmann Series in Database Management Systems. The editors of this text have met their main goal of providing the first textbook integrating knowledge discovery, data mining, and visualization. Although it contributes greatly to our under- standing of the development and current state of the field, a major weakness of this text is that there is no concluding chapter to discuss the contributions of the sum of these contributed papers or give direction to possible future areas of research. "Integration of expertise between two different disciplines is a difficult process of communication and reeducation. Integrating data mining and visualization is particularly complex because each of these fields in itself must draw an a wide range of research experience" (p. 300). Although this work contributes to the crossdisciplinary communication needed to advance visualization in KDD, a more formal call for an interdisciplinary research agenda in a concluding chapter would have provided a more satisfying conclusion to a very good introductory text.
With contributors almost exclusively from the computer science field, the intended audience of this work is heavily slanted towards a computer science perspective. However, it is highly readable and provides introductory material that would be useful to information scientists from a variety of domains. Yet, much interesting work in information visualization from other fields could have been included giving the work more of an interdisciplinary perspective to complement their goals of integrating work in this area. Unfortunately, many of the application chapters are these, shallow, and lack complementary illustrations of visualization techniques or user interfaces used. However, they do provide insight into the many applications being developed in this rapidly expanding field. The authors have successfully put together a highly useful reference text for the data mining and information visualization communities. Those interested in a good introduction and overview of complementary research areas in these fields will be satisfied with this collection of papers. The focus upon integrating data visualization with data mining complements texts in each of these fields, such as Advances in Knowledge Discovery and Data Mining (Fayyad et al., MIT Press) and Readings in Information Visualization: Using Vision to Think (Card et. al., Morgan Kauffman). This unique work is a good starting point for future interaction between researchers in the fields of data visualization and data mining and makes a good accompaniment for a course focused an integrating these areas or to the main reference texts in these fields."

LCSH

Data mining

RSWK

Visualisierung / Computergraphik / Data Mining
Data Mining / Visualisierung / Aufsatzsammlung (BVB)

Subject

Visualisierung / Computergraphik / Data Mining
Data Mining / Visualisierung / Aufsatzsammlung (BVB)
Data mining

Theme

Data Mining

Handbuch Web Mining im Marketing : Konzepte, Systeme, Fallstudien (2002) 0.09

0.088207915 = product of:
  0.17641583 = sum of:
    0.17641583 = product of:
      0.35283166 = sum of:
        0.35283166 = weight(_text_:mining in 6106) [ClassicSimilarity], result of:
          0.35283166 = score(doc=6106,freq=4.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            1.2342855 = fieldWeight in 6106, product of:
              2.0 = tf(freq=4.0), with freq of:
                4.0 = termFreq=4.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.109375 = fieldNorm(doc=6106)
      0.5 = coord(1/2)
  0.5 = coord(1/2)

Theme: Data Mining

Handbuch der Künstlichen Intelligenz (2003) 0.09

0.08639653 = product of:
  0.17279306 = sum of:
    0.17279306 = sum of:
      0.12474483 = weight(_text_:mining in 2916) [ClassicSimilarity], result of:
        0.12474483 = score(doc=2916,freq=2.0), product of:
          0.28585905 = queryWeight, product of:
            5.642448 = idf(docFreq=425, maxDocs=44218)
            0.05066224 = queryNorm
          0.4363858 = fieldWeight in 2916, product of:
            1.4142135 = tf(freq=2.0), with freq of:
              2.0 = termFreq=2.0
            5.642448 = idf(docFreq=425, maxDocs=44218)
            0.0546875 = fieldNorm(doc=2916)
      0.048048235 = weight(_text_:22 in 2916) [ClassicSimilarity], result of:
        0.048048235 = score(doc=2916,freq=2.0), product of:
          0.17741053 = queryWeight, product of:
            3.5018296 = idf(docFreq=3622, maxDocs=44218)
            0.05066224 = queryNorm
          0.2708308 = fieldWeight in 2916, product of:
            1.4142135 = tf(freq=2.0), with freq of:
              2.0 = termFreq=2.0
            3.5018296 = idf(docFreq=3622, maxDocs=44218)
            0.0546875 = fieldNorm(doc=2916)
  0.5 = coord(1/2)

Abstract: Das Handbuch der Künstlichen Intelligenz bietet die umfassendste deutschsprachige Übersicht über die Disziplin "Künstliche Intelligenz". Es vereinigt einführende und weiterführende Beiträge u.a. zu folgenden Themen: - Kognition - Neuronale Netze - Suche, Constraints - Wissensrepräsentation - Logik und automatisches Beweisen - Unsicheres und vages Wissen - Wissen über Raum und Zeit - Fallbasiertes Schließen und modellbasierte Systeme - Planen - Maschinelles Lernen und Data Mining - Sprachverarbeitung - Bildverstehen - Robotik - Software-Agenten Das Handbuch bietet eine moderne Einführung in die Künstliche Intelligenz und zugleich einen aktuellen Überblick über Theorien, Methoden und Anwendungen.
Date: 21. 3.2008 19:10:22

Data Mining im praktischen Einsatz : Verfahren und Anwendungsfälle für Marketing, Vertrieb, Controlling und Kundenunterstützung (2000) 0.08

0.075606786 = product of:
  0.15121357 = sum of:
    0.15121357 = product of:
      0.30242714 = sum of:
        0.30242714 = weight(_text_:mining in 3425) [ClassicSimilarity], result of:
          0.30242714 = score(doc=3425,freq=4.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            1.057959 = fieldWeight in 3425, product of:
              2.0 = tf(freq=4.0), with freq of:
                4.0 = termFreq=4.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.09375 = fieldNorm(doc=3425)
      0.5 = coord(1/2)
  0.5 = coord(1/2)

Theme: Data Mining

Witten, I.H.; Frank, E.: Data Mining : Praktische Werkzeuge und Techniken für das maschinelle Lernen (2000) 0.08

0.075606786 = product of:
  0.15121357 = sum of:
    0.15121357 = product of:
      0.30242714 = sum of:
        0.30242714 = weight(_text_:mining in 6833) [ClassicSimilarity], result of:
          0.30242714 = score(doc=6833,freq=4.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            1.057959 = fieldWeight in 6833, product of:
              2.0 = tf(freq=4.0), with freq of:
                4.0 = termFreq=4.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.09375 = fieldNorm(doc=6833)
      0.5 = coord(1/2)
  0.5 = coord(1/2)

Theme: Data Mining

Intelligent information processing and web mining : Proceedings of the International IIS: IIPWM'03 Conference held in Zakopane, Poland, June 2-5, 2003 (2003) 0.08

0.075606786 = product of:
  0.15121357 = sum of:
    0.15121357 = product of:
      0.30242714 = sum of:
        0.30242714 = weight(_text_:mining in 4642) [ClassicSimilarity], result of:
          0.30242714 = score(doc=4642,freq=4.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            1.057959 = fieldWeight in 4642, product of:
              2.0 = tf(freq=4.0), with freq of:
                4.0 = termFreq=4.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.09375 = fieldNorm(doc=4642)
      0.5 = coord(1/2)
  0.5 = coord(1/2)

Theme: Data Mining

Relational data mining (2001) 0.07
```
0.07072368 = product of:
  0.14144737 = sum of:
    0.14144737 = product of:
      0.28289473 = sum of:
        0.28289473 = weight(_text_:mining in 1303) [ClassicSimilarity], result of:
          0.28289473 = score(doc=1303,freq=14.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            0.9896301 = fieldWeight in 1303, product of:
              3.7416575 = tf(freq=14.0), with freq of:
                14.0 = termFreq=14.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.046875 = fieldNorm(doc=1303)
      0.5 = coord(1/2)
  0.5 = coord(1/2)
```
Abstract

As the first book devoted to relational data mining, this coherently written multi-author monograph provides a thorough introduction and systematic overview of the area. The ferst part introduces the reader to the basics and principles of classical knowledge discovery in databases and inductive logic programmeng; subsequent chapters by leading experts assess the techniques in relational data mining in a principled and comprehensive way; finally, three chapters deal with advanced applications in various fields and refer the reader to resources for relational data mining. This book will become a valuable source of reference for R&D professionals active in relational data mining. Students as well as IT professionals and ambitioned practitioners interested in learning about relational data mining will appreciate the book as a useful text and gentle introduction to this exciting new field.

Theme

Data Mining
Survey of text mining : clustering, classification, and retrieval (2004) 0.07
```
0.06682759 = product of:
  0.13365518 = sum of:
    0.13365518 = product of:
      0.26731035 = sum of:
        0.26731035 = weight(_text_:mining in 804) [ClassicSimilarity], result of:
          0.26731035 = score(doc=804,freq=18.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            0.9351125 = fieldWeight in 804, product of:
              4.2426405 = tf(freq=18.0), with freq of:
                18.0 = termFreq=18.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.0390625 = fieldNorm(doc=804)
      0.5 = coord(1/2)
  0.5 = coord(1/2)
```
Abstract

Extracting content from text continues to be an important research problem for information processing and management. Approaches to capture the semantics of text-based document collections may be based on Bayesian models, probability theory, vector space models, statistical models, or even graph theory. As the volume of digitized textual media continues to grow, so does the need for designing robust, scalable indexing and search strategies (software) to meet a variety of user needs. Knowledge extraction or creation from text requires systematic yet reliable processing that can be codified and adapted for changing needs and environments. This book will draw upon experts in both academia and industry to recommend practical approaches to the purification, indexing, and mining of textual information. It will address document identification, clustering and categorizing documents, cleaning text, and visualizing semantic models of text.

LCSH

Data mining ; Information retrieval
Data mining / Congresses (GBV)

RSWK

Text Mining / Aufsatzsammlung

Subject

Text Mining / Aufsatzsammlung
Data mining ; Information retrieval
Data mining / Congresses (GBV)

Theme

Data Mining
Kantardzic, M.: Data mining : concepts, models, methods, and algorithms (2003) 0.06
```
0.061732687 = product of:
  0.123465374 = sum of:
    0.123465374 = product of:
      0.24693075 = sum of:
        0.24693075 = weight(_text_:mining in 2291) [ClassicSimilarity], result of:
          0.24693075 = score(doc=2291,freq=24.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            0.86381996 = fieldWeight in 2291, product of:
              4.8989797 = tf(freq=24.0), with freq of:
                24.0 = termFreq=24.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.03125 = fieldNorm(doc=2291)
      0.5 = coord(1/2)
  0.5 = coord(1/2)
```
Abstract

This book offers a comprehensive introduction to the exploding field of data mining. We are surrounded by data, numerical and otherwise, which must be analyzed and processed to convert it into information that informs, instructs, answers, or otherwise aids understanding and decision-making. Due to the ever-increasing complexity and size of today's data sets, a new term, data mining, was created to describe the indirect, automatic data analysis techniques that utilize more complex and sophisticated tools than those which analysts used in the past to do mere data analysis. "Data Mining: Concepts, Models, Methods, and Algorithms" discusses data mining principles and then describes representative state-of-the-art methods and algorithms originating from different disciplines such as statistics, machine learning, neural networks, fuzzy logic, and evolutionary computation. Detailed algorithms are provided with necessary explanations and illustrative examples. This text offers guidance: how and when to use a particular software tool (with their companion data sets) from among the hundreds offered when faced with a data set to mine. This allows analysts to create and perform their own data mining experiments using their knowledge of the methodologies and techniques provided. This book emphasizes the selection of appropriate methodologies and data analysis software, as well as parameter tuning. These critically important, qualitative decisions can only be made with the deeper understanding of parameter meaning and its role in the technique that is offered here. Data mining is an exploding field and this book offers much-needed guidance to selecting among the numerous analysis programs that are available.

LCSH

Data mining

RSWK

Data Mining / Lehrbuch

Subject

Data Mining / Lehrbuch
Data mining

Theme

Data Mining
Heyer, G.; Quasthoff, U.; Wittig, T.: Text Mining : Wissensrohstoff Text. Konzepte, Algorithmen, Ergebnisse (2006) 0.06
```
0.056353975 = product of:
  0.11270795 = sum of:
    0.11270795 = product of:
      0.2254159 = sum of:
        0.2254159 = weight(_text_:mining in 5218) [ClassicSimilarity], result of:
          0.2254159 = score(doc=5218,freq=20.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            0.7885561 = fieldWeight in 5218, product of:
              4.472136 = tf(freq=20.0), with freq of:
                20.0 = termFreq=20.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.03125 = fieldNorm(doc=5218)
      0.5 = coord(1/2)
  0.5 = coord(1/2)
```
Abstract

Ein großer Teil des Weltwissens befindet sich in Form digitaler Texte im Internet oder in Intranets. Heutige Suchmaschinen nutzen diesen Wissensrohstoff nur rudimentär: Sie können semantische Zusammen-hänge nur bedingt erkennen. Alle warten auf das semantische Web, in dem die Ersteller von Text selbst die Semantik einfügen. Das wird aber noch lange dauern. Es gibt jedoch eine Technologie, die es bereits heute ermöglicht semantische Zusammenhänge in Rohtexten zu analysieren und aufzubereiten. Das Forschungsgebiet "Text Mining" ermöglicht es mit Hilfe statistischer und musterbasierter Verfahren, Wissen aus Texten zu extrahieren, zu verarbeiten und zu nutzen. Hier wird die Basis für die Suchmaschinen der Zukunft gelegt. Das erste deutsche Lehrbuch zu einer bahnbrechenden Technologie: Text Mining: Wissensrohstoff Text Konzepte, Algorithmen, Ergebnisse Ein großer Teil des Weltwissens befindet sich in Form digitaler Texte im Internet oder in Intranets. Heutige Suchmaschinen nutzen diesen Wissensrohstoff nur rudimentär: Sie können semantische Zusammen-hänge nur bedingt erkennen. Alle warten auf das semantische Web, in dem die Ersteller von Text selbst die Semantik einfügen. Das wird aber noch lange dauern. Es gibt jedoch eine Technologie, die es bereits heute ermöglicht semantische Zusammenhänge in Rohtexten zu analysieren und aufzubereiten. Das For-schungsgebiet "Text Mining" ermöglicht es mit Hilfe statistischer und musterbasierter Verfahren, Wissen aus Texten zu extrahieren, zu verarbeiten und zu nutzen. Hier wird die Basis für die Suchmaschinen der Zukunft gelegt. Was fällt Ihnen bei dem Wort "Stich" ein? Die einen denken an Tennis, die anderen an Skat. Die verschiedenen Zusammenhänge können durch Text Mining automatisch ermittelt und in Form von Wortnetzen dargestellt werden. Welche Begriffe stehen am häufigsten links und rechts vom Wort "Festplatte"? Welche Wortformen und Eigennamen treten seit 2001 neu in der deutschen Sprache auf? Text Mining beantwortet diese und viele weitere Fragen. Tauchen Sie mit diesem Lehrbuch ein in eine neue, faszinierende Wissenschaftsdisziplin und entdecken Sie neue, bisher unbekannte Zusammenhänge und Sichtweisen. Sehen Sie, wie aus dem Wissensrohstoff Text Wissen wird! Dieses Lehrbuch richtet sich sowohl an Studierende als auch an Praktiker mit einem fachlichen Schwerpunkt in der Informatik, Wirtschaftsinformatik und/oder Linguistik, die sich über die Grundlagen, Verfahren und Anwendungen des Text Mining informieren möchten und Anregungen für die Implementierung eigener Anwendungen suchen. Es basiert auf Arbeiten, die während der letzten Jahre an der Abteilung Automatische Sprachverarbeitung am Institut für Informatik der Universität Leipzig unter Leitung von Prof. Dr. Heyer entstanden sind. Eine Fülle praktischer Beispiele von Text Mining-Konzepten und -Algorithmen verhelfen dem Leser zu einem umfassenden, aber auch detaillierten Verständnis der Grundlagen und Anwendungen des Text Mining. Folgende Themen werden behandelt: Wissen und Text Grundlagen der Bedeutungsanalyse Textdatenbanken Sprachstatistik Clustering Musteranalyse Hybride Verfahren Beispielanwendungen Anhänge: Statistik und linguistische Grundlagen 360 Seiten, 54 Abb., 58 Tabellen und 95 Glossarbegriffe Mit kostenlosen e-learning-Kurs "Schnelleinstieg: Sprachstatistik" Zusätzlich zum Buch gibt es in Kürze einen Online-Zertifikats-Kurs mit Mentor- und Tutorunterstützung.

Theme

Data Mining

Knowledge management in fuzzy databases (2000) 0.05

0.054016102 = product of:
  0.108032204 = sum of:
    0.108032204 = product of:
      0.21606441 = sum of:
        0.21606441 = weight(_text_:mining in 4260) [ClassicSimilarity], result of:
          0.21606441 = score(doc=4260,freq=6.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            0.75584245 = fieldWeight in 4260, product of:
              2.4494898 = tf(freq=6.0), with freq of:
                6.0 = termFreq=6.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.0546875 = fieldNorm(doc=4260)
      0.5 = coord(1/2)
  0.5 = coord(1/2)

Abstract: The volume presents recent developments in the introduction of fuzzy, probabilistic and rough elements into basic components of fuzzy databases, and their use (notably querying and information retrieval), from the point of view of data mining and knowledge discovery. The main novel aspect of the volume is that issues related to the use of fuzzy elements in databases, database querying, information retrieval, etc. are presented and discussed from the point of view, and for the purpose of data mining and knowledge discovery that are 'hot topics' in recent years
Theme: Data Mining

Ester, M.; Sander, J.: Knowledge discovery in databases : Techniken und Anwendungen (2000) 0.05

0.050404526 = product of:
  0.10080905 = sum of:
    0.10080905 = product of:
      0.2016181 = sum of:
        0.2016181 = weight(_text_:mining in 1374) [ClassicSimilarity], result of:
          0.2016181 = score(doc=1374,freq=4.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            0.705306 = fieldWeight in 1374, product of:
              2.0 = tf(freq=4.0), with freq of:
                4.0 = termFreq=4.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.0625 = fieldNorm(doc=1374)
      0.5 = coord(1/2)
  0.5 = coord(1/2)

Content: Einleitung.- Statistik- und Datenbank-Grundlagen.Klassifikation.- Assoziationsregeln.- Generalisierung und Data Cubes.- Spatial-, Text-, Web-, Temporal-Data Mining. Ausblick.
Theme: Data Mining

Pang, B.; Lee, L.: Opinion mining and sentiment analysis (2008) 0.05
```
0.050404526 = product of:
  0.10080905 = sum of:
    0.10080905 = product of:
      0.2016181 = sum of:
        0.2016181 = weight(_text_:mining in 1171) [ClassicSimilarity], result of:
          0.2016181 = score(doc=1171,freq=16.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            0.705306 = fieldWeight in 1171, product of:
              4.0 = tf(freq=16.0), with freq of:
                16.0 = termFreq=16.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.03125 = fieldNorm(doc=1171)
      0.5 = coord(1/2)
  0.5 = coord(1/2)
```
Abstract

An important part of our information-gathering behavior has always been to find out what other people think. With the growing availability and popularity of opinion-rich resources such as online review sites and personal blogs, new opportunities and challenges arise as people can, and do, actively use information technologies to seek out and understand the opinions of others. The sudden eruption of activity in the area of opinion mining and sentiment analysis, which deals with the computational treatment of opinion, sentiment, and subjectivity in text, has thus occurred at least in part as a direct response to the surge of interest in new systems that deal directly with opinions as a first-class object. Opinion Mining and Sentiment Analysis covers techniques and approaches that promise to directly enable opinion-oriented information-seeking systems. The focus is on methods that seek to address the new challenges raised by sentiment-aware applications, as compared to those that are already present in more traditional fact-based analysis. The survey includes an enumeration of the various applications, a look at general challenges and discusses categorization, extraction and summarization. Finally, it moves beyond just the technical issues, devoting significant attention to the broader implications that the development of opinion-oriented information-access services have: questions of privacy, vulnerability to manipulation, and whether or not reviews can have measurable economic impact. To facilitate future work, a discussion of available resources, benchmark datasets, and evaluation campaigns is also provided. Opinion Mining and Sentiment Analysis is the first such comprehensive survey of this vibrant and important research area and will be of interest to anyone with an interest in opinion-oriented information-seeking systems.

RSWK

World Wide Web / Meinungsäußerung / Data Mining
Data Mining / Psycholinguistik (BVB)

Subject

World Wide Web / Meinungsäußerung / Data Mining
Data Mining / Psycholinguistik (BVB)

Classification, automation, and new media : Proceedings of the 24th Annual Conference of the Gesellschaft für Klassifikation e.V., University of Passau, March 15 - 17, 2000 (2002) 0.05

0.049810346 = product of:
  0.09962069 = sum of:
    0.09962069 = product of:
      0.19924138 = sum of:
        0.19924138 = weight(_text_:mining in 5997) [ClassicSimilarity], result of:
          0.19924138 = score(doc=5997,freq=10.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            0.6969917 = fieldWeight in 5997, product of:
              3.1622777 = tf(freq=10.0), with freq of:
                10.0 = termFreq=10.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.0390625 = fieldNorm(doc=5997)
      0.5 = coord(1/2)
  0.5 = coord(1/2)

Content: Data Analysis, Statistics, and Classification.- Pattern Recognition and Automation.- Data Mining, Information Processing, and Automation.- New Media, Web Mining, and Automation.- Applications in Management Science, Finance, and Marketing.- Applications in Medicine, Biology, Archaeology, and Others.- Author Index.- Subject Index.
RSWK: Data Mining / Kongress / Passau <2000>
Subject: Data Mining / Kongress / Passau <2000>
Theme: Data Mining

Chakrabarti, S.: Mining the Web : discovering knowledge from hypertext data (2003) 0.05
```
0.047149118 = product of:
  0.094298236 = sum of:
    0.094298236 = product of:
      0.18859647 = sum of:
        0.18859647 = weight(_text_:mining in 2222) [ClassicSimilarity], result of:
          0.18859647 = score(doc=2222,freq=14.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            0.6597534 = fieldWeight in 2222, product of:
              3.7416575 = tf(freq=14.0), with freq of:
                14.0 = termFreq=14.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.03125 = fieldNorm(doc=2222)
      0.5 = coord(1/2)
  0.5 = coord(1/2)
```
Footnote

Rez. in: JASIST 55(2004) no.3, S.275-276 (C. Chen): "This is a book about finding significant statistical patterns on the Web - in particular, patterns that are associated with hypertext documents, topics, hyperlinks, and queries. The term pattern in this book refers to dependencies among such items. On the one hand, the Web contains useful information an just about every topic under the sun. On the other hand, just like searching for a needle in a haystack, one would need powerful tools to locate useful information an the vast land of the Web. Soumen Chakrabarti's book focuses an a wide range of techniques for machine learning and data mining an the Web. The goal of the book is to provide both the technical Background and tools and tricks of the trade of Web content mining. Much of the technical content reflects the state of the art between 1995 and 2002. The targeted audience is researchers and innovative developers in this area, as well as newcomers who intend to enter this area. The book begins with an introduction chapter. The introduction chapter explains fundamental concepts such as crawling and indexing as well as clustering and classification. The remaining eight chapters are organized into three parts: i) infrastructure, ii) learning and iii) applications.
Part I, Infrastructure, has two chapters: Chapter 2 on crawling the Web and Chapter 3 an Web search and information retrieval. The second part of the book, containing chapters 4, 5, and 6, is the centerpiece. This part specifically focuses an machine learning in the context of hypertext. Part III is a collection of applications that utilize the techniques described in earlier chapters. Chapter 7 is an social network analysis. Chapter 8 is an resource discovery. Chapter 9 is an the future of Web mining. Overall, this is a valuable reference book for researchers and developers in the field of Web mining. It should be particularly useful for those who would like to design and probably code their own Computer programs out of the equations and pseudocodes an most of the pages. For a student, the most valuable feature of the book is perhaps the formal and consistent treatments of concepts across the board. For what is behind and beyond the technical details, one has to either dig deeper into the bibliographic notes at the end of each chapter, or resort to more in-depth analysis of relevant subjects in the literature. lf you are looking for successful stories about Web mining or hard-way-learned lessons of failures, this is not the book."

Theme

Data Mining
Witschel, H.F.: Terminologie-Extraktion : Möglichkeiten der Kombination statistischer uns musterbasierter Verfahren (2004) 0.04
```
0.044551726 = product of:
  0.08910345 = sum of:
    0.08910345 = product of:
      0.1782069 = sum of:
        0.1782069 = weight(_text_:mining in 123) [ClassicSimilarity], result of:
          0.1782069 = score(doc=123,freq=8.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            0.6234083 = fieldWeight in 123, product of:
              2.828427 = tf(freq=8.0), with freq of:
                8.0 = termFreq=8.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.0390625 = fieldNorm(doc=123)
      0.5 = coord(1/2)
  0.5 = coord(1/2)
```
Abstract

Die Suche nach Informationen in unstrukturierten natürlichsprachlichen Daten ist Gegenstand des sogenannten Text Mining. In dieser Arbeit wird ein Teilgebiet des Text Mining beleuchtet, nämlich die Extraktion domänenspezifischer Fachbegriffe aus Fachtexten der jeweiligen Domäne. Wofür überhaupt Terminologie-Extraktion? Die Antwort darauf ist einfach: der Schlüssel zum Verständnis vieler Fachgebiete liegt in der Kenntnis der zugehörigen Terminologie. Natürlich genügt es nicht, nur eine Liste der Fachtermini einer Domäne zu kennen, um diese zu durchdringen. Eine solche Liste ist aber eine wichtige Voraussetzung für die Erstellung von Fachwörterbüchern (man denke z.B. an Nachschlagewerke wie das klinische Wörterbuch "Pschyrembel"): zunächst muß geklärt werden, welche Begriffe in das Wörterbuch aufgenommen werden sollen, bevor man sich Gedanken um die genaue Definition der einzelnen Termini machen kann. Ein Fachwörterbuch sollte genau diejenigen Begriffe einer Domäne beinhalten, welche Gegenstand der Forschung in diesem Gebiet sind oder waren. Was liegt also näher, als entsprechende Fachliteratur zu betrachten und das darin enthaltene Wissen in Form von Fachtermini zu extrahieren? Darüberhinaus sind weitere Anwendungen der Terminologie-Extraktion denkbar, wie z.B. die automatische Beschlagwortung von Texten oder die Erstellung sogenannter Topic Maps, welche wichtige Begriffe zu einem Thema darstellt und in Beziehung setzt. Es muß also zunächst die Frage geklärt werden, was Terminologie eigentlich ist, vor allem aber werden verschiedene Methoden entwickelt, welche die Eigenschaften von Fachtermini ausnutzen, um diese aufzufinden. Die Verfahren werden aus den linguistischen und 'statistischen' Charakteristika von Fachbegriffen hergeleitet und auf geeignete Weise kombiniert.

RSWK

Sachtext / Text Mining

Subject

Sachtext / Text Mining

Aberer, K. et al.: ¬The Semantic Web : 6th International Semantic Web Conference, 2nd Asian Semantic Web Conference, ISWC 2007 + ASWC 2007, Busan, Korea, November 11-15, 2007 : proceedings (2007) 0.04

0.0436516 = product of:
  0.0873032 = sum of:
    0.0873032 = product of:
      0.1746064 = sum of:
        0.1746064 = weight(_text_:mining in 2477) [ClassicSimilarity], result of:
          0.1746064 = score(doc=2477,freq=12.0), product of:
            0.28585905 = queryWeight, product of:
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.05066224 = queryNorm
            0.6108129 = fieldWeight in 2477, product of:
              3.4641016 = tf(freq=12.0), with freq of:
                12.0 = termFreq=12.0
              5.642448 = idf(docFreq=425, maxDocs=44218)
              0.03125 = fieldNorm(doc=2477)
      0.5 = coord(1/2)
  0.5 = coord(1/2)

LCSH: Data mining
Data Mining and Knowledge Discovery
RSWK: Semantic Web / Metadatenmodell / Data Mining / Ontologie <Wissensverarbeitung> / Kongress / Pusan <2007> (BVB)
Subject: Data mining
Data Mining and Knowledge Discovery
Semantic Web / Metadatenmodell / Data Mining / Ontologie <Wissensverarbeitung> / Kongress / Pusan <2007> (BVB)

Wissensorganisation und Edutainment : Wissen im Spannungsfeld von Gesellschaft, Gestaltung und Industrie. Proceedings der 7. Tagung der Deutschen Sektion der Internationalen Gesellschaft für Wissensorganisation, Berlin, 21.-23.3.2001 (2004) 0.04
```
0.037027087 = product of:
  0.074054174 = sum of:
    0.074054174 = sum of:
      0.053462073 = weight(_text_:mining in 1442) [ClassicSimilarity], result of:
        0.053462073 = score(doc=1442,freq=2.0), product of:
          0.28585905 = queryWeight, product of:
            5.642448 = idf(docFreq=425, maxDocs=44218)
            0.05066224 = queryNorm
          0.18702249 = fieldWeight in 1442, product of:
            1.4142135 = tf(freq=2.0), with freq of:
              2.0 = termFreq=2.0
            5.642448 = idf(docFreq=425, maxDocs=44218)
            0.0234375 = fieldNorm(doc=1442)
      0.0205921 = weight(_text_:22 in 1442) [ClassicSimilarity], result of:
        0.0205921 = score(doc=1442,freq=2.0), product of:
          0.17741053 = queryWeight, product of:
            3.5018296 = idf(docFreq=3622, maxDocs=44218)
            0.05066224 = queryNorm
          0.116070345 = fieldWeight in 1442, product of:
            1.4142135 = tf(freq=2.0), with freq of:
              2.0 = termFreq=2.0
            3.5018296 = idf(docFreq=3622, maxDocs=44218)
            0.0234375 = fieldNorm(doc=1442)
  0.5 = coord(1/2)
```
Content

Enthält die Beiträge: 1. Wissensgesellschaft Michael NIEHAUS: Durch ein Meer von Unwägbarkeiten - Metaphorik in der Wissensgesellschaft S.3 Karsten WEBER: Aufgaben für eine (globale) Wissensgesellschaft oder "Welcome to the new IT? S.9 Katy TEUBENER: Chronos & Kairos. Inhaltsorganisation und Zeitkultur im Internet S.22 Klaus KRAEMER: Wissen und Nachhaltigkeit. Wissensasymmetrien als Problem einer nachhaltigen Entwicklung S.30 2. Lehre und Lernen Gehard BUDIN: Wissensorganisation als Gestaltungsprinzip virtuellen Lernens - epistemische, kommunikative und methodische Anforderungen S.39 Christan SWERTZ: Webdidaktik: Effiziente Inhaltsproduktion für netzbasierte Trainings S.49 Ingrid LOHMANN: Cognitive Mapping im Cyberpunk - Uber Postmoderne und die Transformation eines für so gut wie tot erklärten Literaturgenres zum Bildungstitel S.54 Rudolf W. KECK, Stefanie KOLLMANN, Christian RITZI: Pictura Paedagogica Online - Konzeption und Verwirklichung S.65 Jadranka LASIC-LASIC, Aida SLAVIC, Mihaela BANEK: Gemeinsame Ausbildung der IT Spezialisten an der Universität Zagreb: Vorteile und Probleme S.76 3. Informationsdesign und Visualisierung Maximilian EIBL, Thomas MANDL: Die Qualität von Visualisierungen: Eine Methode zum Vergleich zweidimensionaler Karten S.89 Udo L. FIGGE: Technische Anleitungen und der Erwerb kohärenten Wissens S.116 Monika WITSCH: Ästhetische Zeichenanalyse - eine Methode zur Analyse fundamentalistischer Agitation im Internet S.123 Oliver GERSTHEIMER, Christian LUPP: Systemdesign - Wissen um den Menschen: Bedürfnisorientierte Produktentwicklung im Mobile Business S.135 Philip ZERWECK: Mehrdimensionale Ordnungssysteme im virtuellen Raum anhand eines Desktops S.141
4. Wissensmanagement und Wissenserschließung René JORNA: Organizational Forms and Knowledge Types S.149 Petra BOSCH-SIJTSERRA: The Virtual Organisation and Knowledge Development: A Case of Expectations S.160 Stefan SMOLNIK, Ludmig NASTARSKY: K-Discovery: Identifikation von verteilten Wissensstrukturen in einer prozessorientierten Groupware-Umgebung S.171 Alexander SIGEL: Wissensorganisation, Topic Maps und Ontology Engineering: Die Verbindung bewährter Begriffsstrukuren mit aktueller XML-Technologie S.185 Gerhard RAHMSTORF: Strukturierung von inhaltlichen Daten: Topic Maps und Concepto S.194 Daniella SAMOWSKI: Informationsdienstleistungen und multimediale Wissensorganisation für die Filmwissenschaft und den Medienstandort Babelsberg oder: Was hat Big Brother mit einer Hochschulbibliothek zu tun? S.207 Harald KLEIN: Web Content Mining S.217 5. Wissensportale Peter HABER, Jan HODEL: Die History Toolbox der Universität Basel S.225 H. Peter OHLY: Gestaltungsprinzipien bei sozialwissenschaftlichen Wissensportalen im Internet S.234 Markus QUANDT: souinet.de - ein Internetjournal mit Berichten aus den Sozialwissenschaften. Ziele und Konzept einer neuen Informationsplattform S.247 Jörn SIEGLERSCHMIDT: Das Museum als Interface S.264

Medien-Informationsmanagement : Archivarische, dokumentarische, betriebswirtschaftliche, rechtliche und Berufsbild-Aspekte ; [Frühjahrstagung der Fachgruppe 7 im Jahr 2000 in Weimar und Folgetagung 2001 in Köln] (2003) 0.04

0.037027087 = product of:
  0.074054174 = sum of:
    0.074054174 = sum of:
      0.053462073 = weight(_text_:mining in 1833) [ClassicSimilarity], result of:
        0.053462073 = score(doc=1833,freq=2.0), product of:
          0.28585905 = queryWeight, product of:
            5.642448 = idf(docFreq=425, maxDocs=44218)
            0.05066224 = queryNorm
          0.18702249 = fieldWeight in 1833, product of:
            1.4142135 = tf(freq=2.0), with freq of:
              2.0 = termFreq=2.0
            5.642448 = idf(docFreq=425, maxDocs=44218)
            0.0234375 = fieldNorm(doc=1833)
      0.0205921 = weight(_text_:22 in 1833) [ClassicSimilarity], result of:
        0.0205921 = score(doc=1833,freq=2.0), product of:
          0.17741053 = queryWeight, product of:
            3.5018296 = idf(docFreq=3622, maxDocs=44218)
            0.05066224 = queryNorm
          0.116070345 = fieldWeight in 1833, product of:
            1.4142135 = tf(freq=2.0), with freq of:
              2.0 = termFreq=2.0
            3.5018296 = idf(docFreq=3622, maxDocs=44218)
            0.0234375 = fieldNorm(doc=1833)
  0.5 = coord(1/2)

Date: 11. 5.2008 19:49:22
Theme: Data Mining

Search (180 results, page 1 of 9)

Authors

Languages

Types

Themes

Subjects

Classifications