Session
The Lucene Scalpel: Tuning Low-Level Segments for Multi-Lingual Concurrency
As European search applications scale, they face a unique challenge: maintaining high-concurrency relevance across multi-lingual datasets. When an index grows to 25GB+, standard Lucene merge policies often lead to "Segment Fragmentation," causing significant latency spikes in k-NN search. This session moves beyond basic vector implementation to perform "surgery" on the underlying Lucene engine. We explore how to tune index merge policies and custom analyzers specifically for multi-lingual shards, ensuring that 60+ concurrent streams remain performant. Attendees will learn how to optimize the physical storage of vector data to reduce I/O overhead and maintain sub-second relevance without bloating hardware costs.
Divyanshu Mishra
Lead Devops Engineer - Eastern Enterprise
Delhi, India
Links
Please note that Sessionize is not responsible for the accuracy or validity of the data provided by speakers. If you suspect this profile to be fake or spam, please let us know.
Jump to top