Introducing Multi-valued Vector Fields in Apache Lucene
Formal Metadata
| Title | Introducing Multi-valued Vector Fields in Apache Lucene |
|
| Title of Series | |
| Number of Parts | 60 |
| Author | |
| License | CC Attribution 3.0 Unported: You are free to use, adapt and copy, distribute and transmit the work or content in adapted or unchanged form for any legal purpose as long as the work is attributed to the author in the manner specified by the author or licensor. |
| Identifiers | |
| Publisher | |
| Release Date | |
| Language | |
Content Metadata
| Subject Area | |
| Genre | |
| Abstract | Since the introduction of native vector-based search in Apache Lucene happened, many features have been developed, but the support for multiple vectors in a dedicated KNN vector field remained to explore.
Having the possibility of indexing (and searching) multiple values per field unlocks the possibility of working with long textual documents, splitting them in paragraphs and encoding each paragraph as a separate vector: scenario that is often encountered by many businesses.
This talk explores the challenges, the technical design and the implementation activities happened during the work for this contribution to the Apache Lucene project.
The audience is expected to get an understanding of how multi-valued fields can work in a vector-based search use-case and how this feature has been implemented. |
|