Apache Solr : a practical approach to enterprise search /

Saved in:
Bibliographic Details
Author / Creator:Shahi, Dikshant, author.
Imprint:[Berkeley, California] : Apress, 2015.
New York, NY : Distributed to the book trade worldwide by Springer,
©2015
Description:1 online resource (xxvi, 299 pages) : color illustrations.
Language:English
Series:The expert's voice in enterprise search
Expert's voice in enterprise search.
Subject:
Format: E-Resource Book
URL for this record:http://pi.lib.uchicago.edu/1001/cat/bib/11097188
Hidden Bibliographic Details
ISBN:9781484210703
1484210700
1484210719 (print)
9781484210710 (print)
Digital file characteristics:text file PDF
Notes:Includes index.
Online resource; title from PDF title page (SpringerLink, viewed January 5, 2016).
Summary:Build an enterprise search engine using Apache Solr: index and search documents; ingest data from varied sources; apply various text processing techniques; utilize different search capabilities; and customize Solr to retrieve the desired results. Apache Solr: A Practical Approach to Enterprise Search explains each essential concept-backed by practical and industry examples--to help you attain expert-level knowledge. The book, which assumes a basic knowledge of Java, starts with an introduction to Solr, followed by steps to setting it up, indexing your first set of documents, and searching them. It then introduces you to information retrieval and its implementation in Apache Solr; this will help you understand your search problem, decide the approach to build an effective solution, and use various metrics to evaluate the results. The book next covers the schema design and techniques to build a text analysis chain for cleansing, normalizing and enriching your documents and addressing different types of search queries. It describes various popular matching techniques which are generally applied to improve the precision and recall of searches. You will learn the end-to-end process of data ingestion from varied sources, metadata extraction, pre-processing and transformation of content, various search components, query parsers and other advanced search capabilities. After covering out-of-the-box features, Solr expert Dikshant Shahi dives into ways you can customize Solr for your business and its specific requirements, along with ways to plug in your own components. Most important, you will learn about implementations for Solr scoring, factors affecting the document score, and tuning the score for the application at hand. The book explains why textual scoring is not sufficient for practical ranking of documents and ways to integrate real-world factors for contributing to the document ranking. You'll see how to influence user experience by providing suggestions and recommendations. You'll also see integration of Solr with important related technologies such as OpenNLP and Tika. Additionally, you will learn about scaling Solr using SolrCloud. This book concludes with coverage of semantic search capabilities, which is crucial for taking the search experience to the next level. By the end of Apache Solr, you will be proficient in designing and developing your search engine. .
Other form:Printed edition: 9781484210710
Standard no.:10.1007/978-1-4842-1070-3
Description
Summary:

Build anenterprise search engine using Apache Solr: index and search documents; ingestdata from varied sources; apply various text processing techniques; utilizedifferent search capabilities; and customize Solr to retrieve the desiredresults. Apache Solr: APractical Approach to Enterprise Search explains each essentialconcept-backed by practical and industry examples--to help you attainexpert-level knowledge.

The book,which assumes a basic knowledge of Java, starts with an introduction to Solr,followed by steps to setting it up, indexing your first set of documents, andsearching them. It then introduces you to information retrieval and itsimplementation in Apache Solr; this will help you understand your searchproblem, decide the approach to build an effective solution, and use variousmetrics to evaluate the results.

The booknext covers the schema design and techniques to build a text analysis chain forcleansing, normalizing and enriching your documents and addressing differenttypes of search queries. It describes various popular matching techniques whichare generally applied to improve the precision and recall of searches.

You willlearn the end-to-end process of data ingestion from varied sources, metadataextraction, pre-processing and transformation of content, various searchcomponents, query parsers and other advanced search capabilities.

Aftercovering out-of-the-box features, Solr expert Dikshant Shahi dives into waysyou can customize Solr for your business and its specific requirements, alongwith ways to plug in your own components. Most important, you will learn aboutimplementations for Solr scoring, factors affecting the document score, andtuning the score for the application at hand. The book explains why textualscoring is not sufficient for practical ranking of documents and ways tointegrate real-world factors for contributing to the document ranking.

You'll seehow to influence user experience by providing suggestions and recommendations.You'll also see integration of Solr with important related technologies such asOpenNLP and Tika. Additionally, you will learn about scaling Solr usingSolrCloud.

This book concludes withcoverage of semantic search capabilities, which is crucial for taking thesearch experience to the next level. By the end of Apache Solr, you will be proficient in designing anddeveloping your search engine.

Item Description:Includes index.
Physical Description:1 online resource (xxvi, 299 pages) : color illustrations.
ISBN:9781484210703
1484210700
1484210719 (print)
9781484210710 (print)