Mining the Web
Title | Mining the Web PDF eBook |
Author | Soumen Chakrabarti |
Publisher | Morgan Kaufmann |
Pages | 366 |
Release | 2002-10-09 |
Genre | Computers |
ISBN | 1558607544 |
The definitive book on mining the Web from the preeminent authority.
Mining the Social Web
Title | Mining the Social Web PDF eBook |
Author | Matthew Russell |
Publisher | "O'Reilly Media, Inc." |
Pages | 356 |
Release | 2011-01-21 |
Genre | Computers |
ISBN | 1449388345 |
Facebook, Twitter, and LinkedIn generate a tremendous amount of valuable social data, but how can you find out who's making connections with social media, what they’re talking about, or where they’re located? This concise and practical book shows you how to answer these questions and more. You'll learn how to combine social web data, analysis techniques, and visualization to help you find what you've been looking for in the social haystack, as well as useful information you didn't know existed. Each standalone chapter introduces techniques for mining data in different areas of the social Web, including blogs and email. All you need to get started is a programming background and a willingness to learn basic Python tools. Get a straightforward synopsis of the social web landscape Use adaptable scripts on GitHub to harvest data from social network APIs such as Twitter, Facebook, and LinkedIn Learn how to employ easy-to-use Python tools to slice and dice the data you collect Explore social connections in microformats with the XHTML Friends Network Apply advanced mining techniques such as TF-IDF, cosine similarity, collocation analysis, document summarization, and clique detection Build interactive visualizations with web technologies based upon HTML5 and JavaScript toolkits "Let Matthew Russell serve as your guide to working with social data sets old (email, blogs) and new (Twitter, LinkedIn, Facebook). Mining the Social Web is a natural successor to Programming Collective Intelligence: a practical, hands-on approach to hacking on data from the social Web with Python." --Jeff Hammerbacher, Chief Scientist, Cloudera "A rich, compact, useful, practical introduction to a galaxy of tools, techniques, and theories for exploring structured and unstructured data." --Alex Martelli, Senior Staff Engineer, Google
Data Mining Your Website
Title | Data Mining Your Website PDF eBook |
Author | Jesus Mena |
Publisher | Digital Press |
Pages | 388 |
Release | 1999-07-15 |
Genre | Business & Economics |
ISBN | 9781555582227 |
Turn Web data into knowledge about your customers. This exciting book will help companies create, capture, enhance, and analyze one of their most valuable new sources of marketing information-usage and transactional data from a website. A company's website is a primary point of contact with its customers and a medium in which visitor's actions are messages about who they are and what they want. Data Mining Your Website will teach you the tools, techniques, and technologies you'll need to profile current and potential customers and predict on-line interests and behavior. You'll learn how to extract from the huge pools of information your website generates, insights into on-line buying patterns, and how to apply this knowledge to design a website that better attracts, engages, and retains on-line customers. Data Mining Your Website explains how data mining is a foundation for the new field of web-based, interactive retailing, marketing, and advertising. This innovative book will help web developers and marketers, webmasters, and data management professionals harness powerful new tools and processes. The first book to apply data mining specifically to e-commerce Learn effective methods for gathering, managing, and mining Web customer information Use data mining to profile customers and create personalized e-commerce programs
Data Mining: Concepts and Techniques
Title | Data Mining: Concepts and Techniques PDF eBook |
Author | Jiawei Han |
Publisher | Elsevier |
Pages | 740 |
Release | 2011-06-09 |
Genre | Computers |
ISBN | 0123814804 |
Data Mining: Concepts and Techniques provides the concepts and techniques in processing gathered data or information, which will be used in various applications. Specifically, it explains data mining and the tools used in discovering knowledge from the collected data. This book is referred as the knowledge discovery from data (KDD). It focuses on the feasibility, usefulness, effectiveness, and scalability of techniques of large data sets. After describing data mining, this edition explains the methods of knowing, preprocessing, processing, and warehousing data. It then presents information about data warehouses, online analytical processing (OLAP), and data cube technology. Then, the methods involved in mining frequent patterns, associations, and correlations for large data sets are described. The book details the methods for data classification and introduces the concepts and methods for data clustering. The remaining chapters discuss the outlier detection and the trends, applications, and research frontiers in data mining. This book is intended for Computer Science students, application developers, business professionals, and researchers who seek information on data mining. - Presents dozens of algorithms and implementation examples, all in pseudo-code and suitable for use in real-world, large-scale data mining projects - Addresses advanced topics such as mining object-relational databases, spatial databases, multimedia databases, time-series databases, text databases, the World Wide Web, and applications in several fields - Provides a comprehensive, practical look at the concepts and techniques you need to get the most out of your data
Web Data Mining
Title | Web Data Mining PDF eBook |
Author | Bing Liu |
Publisher | Springer Science & Business Media |
Pages | 637 |
Release | 2011-06-25 |
Genre | Computers |
ISBN | 3642194605 |
Liu has written a comprehensive text on Web mining, which consists of two parts. The first part covers the data mining and machine learning foundations, where all the essential concepts and algorithms of data mining and machine learning are presented. The second part covers the key topics of Web mining, where Web crawling, search, social network analysis, structured data extraction, information integration, opinion mining and sentiment analysis, Web usage mining, query log mining, computational advertising, and recommender systems are all treated both in breadth and in depth. His book thus brings all the related concepts and algorithms together to form an authoritative and coherent text. The book offers a rich blend of theory and practice. It is suitable for students, researchers and practitioners interested in Web mining and data mining both as a learning text and as a reference book. Professors can readily use it for classes on data mining, Web mining, and text mining. Additional teaching materials such as lecture slides, datasets, and implemented algorithms are available online.
Data Mining with R
Title | Data Mining with R PDF eBook |
Author | Luis Torgo |
Publisher | CRC Press |
Pages | 426 |
Release | 2016-11-30 |
Genre | Business & Economics |
ISBN | 1315399091 |
Data Mining with R: Learning with Case Studies, Second Edition uses practical examples to illustrate the power of R and data mining. Providing an extensive update to the best-selling first edition, this new edition is divided into two parts. The first part will feature introductory material, including a new chapter that provides an introduction to data mining, to complement the already existing introduction to R. The second part includes case studies, and the new edition strongly revises the R code of the case studies making it more up-to-date with recent packages that have emerged in R. The book does not assume any prior knowledge about R. Readers who are new to R and data mining should be able to follow the case studies, and they are designed to be self-contained so the reader can start anywhere in the document. The book is accompanied by a set of freely available R source files that can be obtained at the book’s web site. These files include all the code used in the case studies, and they facilitate the "do-it-yourself" approach followed in the book. Designed for users of data analysis tools, as well as researchers and developers, the book should be useful for anyone interested in entering the "world" of R and data mining. About the Author Luís Torgo is an associate professor in the Department of Computer Science at the University of Porto in Portugal. He teaches Data Mining in R in the NYU Stern School of Business’ MS in Business Analytics program. An active researcher in machine learning and data mining for more than 20 years, Dr. Torgo is also a researcher in the Laboratory of Artificial Intelligence and Data Analysis (LIAAD) of INESC Porto LA.
Data Mining
Title | Data Mining PDF eBook |
Author | Ian H. Witten |
Publisher | Elsevier |
Pages | 665 |
Release | 2011-02-03 |
Genre | Computers |
ISBN | 0080890369 |
Data Mining: Practical Machine Learning Tools and Techniques, Third Edition, offers a thorough grounding in machine learning concepts as well as practical advice on applying machine learning tools and techniques in real-world data mining situations. This highly anticipated third edition of the most acclaimed work on data mining and machine learning will teach you everything you need to know about preparing inputs, interpreting outputs, evaluating results, and the algorithmic methods at the heart of successful data mining. Thorough updates reflect the technical changes and modernizations that have taken place in the field since the last edition, including new material on Data Transformations, Ensemble Learning, Massive Data Sets, Multi-instance Learning, plus a new version of the popular Weka machine learning software developed by the authors. Witten, Frank, and Hall include both tried-and-true techniques of today as well as methods at the leading edge of contemporary research. The book is targeted at information systems practitioners, programmers, consultants, developers, information technology managers, specification writers, data analysts, data modelers, database R&D professionals, data warehouse engineers, data mining professionals. The book will also be useful for professors and students of upper-level undergraduate and graduate-level data mining and machine learning courses who want to incorporate data mining as part of their data management knowledge base and expertise. - Provides a thorough grounding in machine learning concepts as well as practical advice on applying the tools and techniques to your data mining projects - Offers concrete tips and techniques for performance improvement that work by transforming the input or output in machine learning methods - Includes downloadable Weka software toolkit, a collection of machine learning algorithms for data mining tasks—in an updated, interactive interface. Algorithms in toolkit cover: data pre-processing, classification, regression, clustering, association rules, visualization