|
|
Customers who bought this book also bought:
Click here for more suggestions...
Look for similar books by subject:
Editorial Reviews
Amazon.com Of all the tasks programmers are asked to perform, storing, compressing, and retrieving information are some of the most challenging--and critical to many applications. Managing Gigabytes: Compressing and Indexing Documents and Images is a treasure trove of theory, practical illustration, and general discussion in this fascinating technical subject. Ian Witten, Alistair Moffat, and Timothy Bell have updated their original work with this even more impressive second edition. This version adds recent techniques such as block-sorting, new indexing techniques, new lossless compression strategies, and many other elements to the mix. In short, this work is a comprehensive summary of text and image compression, indexing, and querying techniques. The history of relevant algorithm development is woven well with a practical discussion of challenges, pitfalls, and specific solutions. This title is a textbook-style exposition on the topic, with its information organized very clearly into topics such as compression, indexing, and so forth. In addition to diagrams and example text transformations, the authors use "pseudo-code" to present algorithms in a language-independent manner wherever possible. They also supplement the reading with mg--their own implementation of the techniques. The mg C language source code is freely available on the Web. Alone, this book is an impressive collection of information. Nevertheless, the authors list numerous titles for further reading in selected topics. Whether you're in the midst of application development and need solutions fast or are merely curious about how top-notch information management is done, this hardcover is an excellent investment. --Stephen W. Plain Topics covered: Text compression models, including Huffman, LZW, and their variants; trends in information management; index creation and compression; image compression; performance issues; and overall system implementation. --This text refers to the hardcover edition of this title From Book News, Inc. , August 1, 1994 The end result of applying the techniques described here is a computer system that can store millions of documents, and retrieve the documents that contain any given combination of keywords in a matter of seconds or fractions of a second. Written for an eclectic audience of information professionals and for graduate courses. Sections for technically or theoretically oriented readers can be skipped by others without loss of continuity. Annotation copyright Book News, Inc. Portland, Or. Steve Kirsch, Cofounder, Infoseek Corporation "This book is the Bible for anyone who needs to manage large data collections. It's required reading for our search gurus at Infoseek. The authors have done an outstanding job of incorporating and describing the most significant new research in information retrieval over the past five years into this second edition." --This text refers to the hardcover edition of this title Michael Lesk, National Science Foundation "The new edition of Witten, Moffat, and Bell not only has newer and better text search algorithms but much material on image analysis and joint image/text processing. If you care about search engines, you need this book: it is the only one with full details of how they work. The book is both detailed and enjoyable; the authors have combined elegant writing with top-grade programming." --This text refers to the hardcover edition of this title read more
See all
8 editorial reviews...
All Customer Reviews
Avg. Customer Review:
Number of Reviews: 7
Write an online review and share your thoughts with other readers!
|
3 of 3 people found the following review helpful:
|
|
Good introduction to searching/indexing in data.
|
December 29, 1999
|
|
|
Reviewer:
Amund Tveit
(see more about me)
from Trondheim, Norway
|
|
|
MG gave a good introduction to the components of practical Information Retrieval (IR). You can clearly see that the authors have a genuine interest in the field! But, I would like some more theoretical analysis of the algorithms used(i.e. O-notation), and more focus on parallell implementations of IR systems. Another book related to the same area worth mentioning is "Modern Information Retrieval".
--This text refers to the Hardcover edition.
|
|
|
|
This is a great book.
|
September 17, 1999
|
|
|
Reviewer:
las@cs.nmhu.edu
from New Mexico, USA
|
|
|
This is one of those rare books that succeeds both on a theoretical and practical level. The theory underlying management and retrieval of large collections of mixed text and image data is thoroughly covered. The authors' experience in developing the accompanying software shines through in the clarity of their explanations and enables them to give practical information regarding the techniques discussed. The software is not just of academic interest, either - an appendix describes a digital library, accessible over the web, that is supported by the mg software. In summary, this is a great book - readable, thorough and practical.
--This text refers to the Hardcover edition.
|
|
|
|
4 of 4 people found the following review helpful:
|
|
Compression, Algorithms, Full Text Retrieval
|
September 14, 1999
|
|
|
Reviewer:
Darryl Lovato (dlovato@aladdinsys.com)
from United States
|
|
|
Managing Gigabytes is a must read for anyone iterested in how to transmit, access, store, and search large amounts of data. I'm the President and CTO of Aladdin Systems, Inc, the creators of the StuffIt compression product line for Mac and Windows, and I find it an invaluable addition to my reference library. The authors take complex information and present it in an organized, easy to read format, suitable for novices to experts. I highly recommend this book.
--This text refers to the Hardcover edition.
|
|
|
|
2 of 2 people found the following review helpful:
|
|
Best text available. Has no competition.
|
September 10, 1999
|
|
|
Reviewer:
A reader
from Texas, USA
|
|
|
This text sets the standard for future information retrieval texts and has replaced the Salton books as the canonical academic text. The second edition is highly readable and contains a thorough updating of the algorithms and data structures in the field. I like the text because of its readability, conciseness, thoroughness, and attention to detail. The comparisons of algorithms on realistic sized collections is unparalleled in other texts. I have used this text for the past 5 years in a graduate level information storage and retrieval class but I believe it has a much wider audience due to the quality of writing. Additionally, the free availability of the mg system which implements many of the best algorithms of the text allows the reader/student to take advantage of the technology without having to start from scratch. Highly recommended.
--This text refers to the Hardcover edition.
|
|
|
See all 7 customer reviews...
Customers who bought titles by Ian H. Witten also bought titles by these authors:
Look for similar books by subject:
|