Managing Gigabytes: Compressing and Indexing Documents and Images, Second Edition
Morgan Kaufmann, 17 mai 1999 - 519 pages
"This book is the Bible for anyone who needs to manage large data collections. It's required reading for our search gurus at Infoseek. The authors have done an outstanding job of incorporating and describing the most significant new research in information retrieval over the past five years into this second edition."
"The new edition of Witten, Moffat, and Bell not only has newer and better text search algorithms but much material on image analysis and joint image/text processing. If you care about search engines, you need this book: it is the only one with full details of how they work. The book is both detailed and enjoyable; the authors have combined elegant writing with top-grade programming."
"The coverage of compression, file organizations, and indexing techniques for full text and document management systems is unsurpassed. Students, researchers, and practitioners will all benefit from reading this book."
In this fully updated second edition of the highly acclaimed Managing Gigabytes, authors Witten, Moffat, and Bell continue to provide unparalleled coverage of state-of-the-art techniques for compressing and indexing data. Whatever your field, if you work with large quantities of information, this book is essential reading--an authoritative theoretical resource and a practical guide to meeting the toughest storage and access challenges. It covers the latest developments in compression and indexing and their application on the Web and in digital libraries. It also details dozens of powerful techniques supported by mg, the authors' own system for compressing, storing, and retrieving text, images, and textual images. mg's source code is freely available on the Web.
Avis des internautes - Rédiger un commentaire
Avis des utilisateurs
LibraryThing ReviewAvis d'utilisateur - juha - LibraryThing
A hard-core approach to information retrieval. I didn't appreaciate this book until recently, when I started to look for ways to reduce I/O. The use of compression in storing the text, integers, lexicon and inverted list is detailed beautifully. Consulter l'avis complet
two Text Compression
five Index Construction
six Image Compression
seven Textual Images
eight Mixed Text and Images
Autres éditions - Tout afficher
Practical Digital Libraries: Books, Bytes, and Bucks
Michael Lesk,MICHAEL AUTOR LESK
Aperçu limité - 1997