Sease at Open Source Summit 2017

June 5, 2017
3 mins read

Open Source Summit connects the open source ecosystem under one roof. It’s a unique environment for cross-collaboration between developers, sysadmins, devops, architects and others who are driving technology forward.

Location: Tokyo (Japan)

Date: 31 May – 2 June 2017

our talk

Being your core domain involving real world entities ( such as hotels, restaurant, cars …) or text documents, searching for similar entities, given one in input, is a very common use case for most of the systems that involve information retrieval. This presentation will start describing how much this problem is present across a variety of different scenarios and how you can use the More Like This feature in the Apache Lucene library to solve it. Building on the introduction the focus will be on how the More Like This module internally works, all the components involved end to end, BM25 text similarity metric and how this has been included through a cospicuos refactor and testing process. The presentation will include real world usage examples and future developments such as improved query building through positional phrase queries and term relevancy scoring pluggability.

our speaker

Alessandro Benedetti

FOUNDER @ SEASE

APACHE LUCENE/SOLR COMMITTER
APACHE SOLR PMC MEMBER

Senior Search Software Engineer, his focus is on R&D in Information Retrieval, Information Extraction, Natural Language Processing, and Machine Learning.
He firmly believes in Open Source as a way to build a bridge between Academia and Industry and facilitate the progress of applied research.

slides

Advanced Document Similarity with Apache Lucene from Sease

open source summit

Other posts you may find useful

Lexically accelerated vector search SeededKnnVectorQuery Support in Apache Solr 10

We are Sease, an Information Retrieval Company based in London, focused on providing R&D project guidance and implementation, Search consulting services, Training, and Search solutions using open source software like Apache Lucene/Solr, Elasticsearch, OpenSearch and Vespa.

Sign up for our Newsletter

Did you like this post? Don’t forget to subscribe to our Newsletter to stay always updated in the Information Retrieval world!

About the company

about our work

Rated Ranking Evaluator
(RRE)

Rated Ranking Evaluator Enterprise (RREE)

Apache Solr LLM Highlighter plugin

News

Main Blog

TIPS AND TRICKS

LATEST BLOG POST

contact us

Don't miss all the news - subscribe to our newsletter!

Sease at Open Source Summit 2017

our talk

our speaker

Alessandro Benedetti

slides

Other posts you may find useful

Lexically accelerated vector search: SeededKnnVectorQuery Support in Apache Solr 10

Solr Is Learning To Rank Better – Part 1 – Data Collection

How to calculate aggregations in Elasticsearch as percentages?

Lisa Biella

Lisa Biella

Follow Us

Top Categories

Recent Posts

Boosted K-Nearest Neighbor Search

Vector Search Doctor (Part 2): Bridging the Gap Between Theory and Practice in Vector Search

Vector Search Doctor (Part 1): Beyond the MTEB Leaderboard for Custom Datasets

Monthly video

Sign up for our Newsletter

Leave a Reply Cancel reply

Quick Links

Services

Subscribe

About the company

about our work

Rated Ranking Evaluator (RRE)

Rated Ranking Evaluator Enterprise (RREE)

Apache Solr LLM Highlighter plugin

News

Main Blog

TIPS AND TRICKS

LATEST BLOG POST

contact us

Don't miss all the news - subscribe to our newsletter!

Sease at Open Source Summit 2017

our talk

our speaker

Alessandro Benedetti

slides

Other posts you may find useful

Lexically accelerated vector search: SeededKnnVectorQuery Support in Apache Solr 10

Solr Is Learning To Rank Better – Part 1 – Data Collection

How to calculate aggregations in Elasticsearch as percentages?

Lisa Biella

Lisa Biella

Follow Us

Top Categories

Recent Posts

Boosted K-Nearest Neighbor Search

Vector Search Doctor (Part 2): Bridging the Gap Between Theory and Practice in Vector Search

Vector Search Doctor (Part 1): Beyond the MTEB Leaderboard for Custom Datasets

Monthly video

Sign up for our Newsletter

Leave a Reply Cancel reply

Rated Ranking Evaluator
(RRE)