Apache Solr Distributed Search Document Frequency Scalability Search Tips And Tricks

Distributed Search Tips for Apache Solr

Distributed search is the foundation for Apache Solr Scalability : It’s possible to distributed search across different Apache Solr nodes of the same collection ( both in a  legacy[1] or SolrCloud[2] architecture), but it is also possible to distribute search across different collections in a SolrCloud cluster. Aggregating results from different collections may be useful…

Apache Solr Data Preparation Feature Engineering Learning To Rank Machine Learning Main Blog RankLib Search Signal Processing

Solr Is Learning To Rank Better – Part 1 – Data Collection

Learning To Rank In Apache Solr Introduction This blog post is about the journey necessary to bring Learning To Rank In Apache Solr search engines. Learning to Rank[1] is the application of Machine Learning in the construction of ranking models for Information Retrieval systems. Introducing supervised learning from user behaviour and signals can improve the relevancy…

Apache Lucene Apache Solr Document Frequency Indexing Indexing options Lucene index Main Blog Norms Solr schema Term Frequency Term offsets Term positions

Exploring Solr Internals : The Lucene Inverted Index

    Introduction This blog post is about the Lucene Inverted Index and how Apache Solr internally works. When playing with Solr systems, understanding and properly configuring the underline Lucene Index is fundamental to deeply control your search. With a better knowledge of how the index looks like and how each component is used, you…

Analysis Apache Lucene Apache Solr Autocomplete Autosuggestion FST Lucene index Main Blog Ngrams Suggester Token filters Tokenizer

Solr : " You complete me! " : The Apache Solr Suggester

This blog post is about the Apache Solr Autocomplete feature. It is clear that the current documentation available on the wiki is not enough to fully understand the Solr Suggester : this blog post will describe all the available implementations with examples and tricks and tips. Introduction If there’s one thing that months of Solr-user…

Apache Lucene Apache Solr Classification Indexing Machine Learning Main Blog Search Update Request Processor

Solr Document Classification – Part 1 – Indexing Time

Introduction This blog post is about the Solr classification module and the way Lucene classification has been integrated at indexing time. In the previous blog [1] we have explored the world of Lucene Classification and the extension to use it for Document Classification . It comes natural to integrate Solr with the Classification module and…

Apache Lucene Apache Solr Classification Machine Learning Main Blog Search Search Library

Lucene Document Classification

Introduction This blog post describes the approach used in the Lucene Classification module to adapt text classification to document ( multi field ) classification. Machine Learning and Search have been always strictly associated. Machine Learning can help to improve the Search Experience in a lot of ways, extracting more information from the corpus of documents,…