Efficient SPLADE model for passage retrieval. This architecture uses two distinct models for query and document inference. This is the
query
one, please also download the
doc
one (
https://huggingface.co/naver/efficient-splade-V-large-doc
). For additional details, please visit:
If you use our checkpoint, please cite our work (need to update):
@inproceedings{10.1145/3477495.3531833,
author = {Lassance, Carlos and Clinchant, St\'{e}phane},
title = {An Efficiency Study for SPLADE Models},
year = {2022},
isbn = {9781450387323},
publisher = {Association for Computing Machinery},
address = {New York, NY, USA},
url = {https://doi.org/10.1145/3477495.3531833},
doi = {10.1145/3477495.3531833},
abstract = {Latency and efficiency issues are often overlooked when evaluating IR models based on Pretrained Language Models (PLMs) in reason of multiple hardware and software testing scenarios. Nevertheless, efficiency is an important part of such systems and should not be overlooked. In this paper, we focus on improving the efficiency of the SPLADE model since it has achieved state-of-the-art zero-shot performance and competitive results on TREC collections. SPLADE efficiency can be controlled via a regularization factor, but solely controlling this regularization has been shown to not be efficient enough. In order to reduce the latency gap between SPLADE and traditional retrieval systems, we propose several techniques including L1 regularization for queries, a separation of document/query encoders, a FLOPS-regularized middle-training, and the use of faster query encoders. Our benchmark demonstrates that we can drastically improve the efficiency of these models while increasing the performance metrics on in-domain data. To our knowledge, we propose the first neural models that, under the same computing constraints, achieve similar latency (less than 4ms difference) as traditional BM25, while having similar performance (less than 10% MRR@10 reduction) as the state-of-the-art single-stage neural rankers on in-domain data.},
booktitle = {Proceedings of the 45th International ACM SIGIR Conference on Research and Development in Information Retrieval},
pages = {2220–2226},
numpages = {7},
keywords = {splade, latency, information retrieval, sparse representations},
location = {Madrid, Spain},
series = {SIGIR '22}
}
Runs of naver efficient-splade-V-large-query on huggingface.co
206
Total runs
0
24-hour runs
-8
3-day runs
-58
7-day runs
-370
30-day runs
More Information About efficient-splade-V-large-query huggingface.co Model
More efficient-splade-V-large-query license Visit here:
efficient-splade-V-large-query huggingface.co is an AI model on huggingface.co that provides efficient-splade-V-large-query's model effect (), which can be used instantly with this naver efficient-splade-V-large-query model. huggingface.co supports a free trial of the efficient-splade-V-large-query model, and also provides paid use of the efficient-splade-V-large-query. Support call efficient-splade-V-large-query model through api, including Node.js, Python, http.
efficient-splade-V-large-query huggingface.co is an online trial and call api platform, which integrates efficient-splade-V-large-query's modeling effects, including api services, and provides a free online trial of efficient-splade-V-large-query, you can try efficient-splade-V-large-query online for free by clicking the link below.
naver efficient-splade-V-large-query online free url in huggingface.co:
efficient-splade-V-large-query is an open source model from GitHub that offers a free installation service, and any user can find efficient-splade-V-large-query on GitHub to install. At the same time, huggingface.co provides the effect of efficient-splade-V-large-query install, users can directly use efficient-splade-V-large-query installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
efficient-splade-V-large-query install url in huggingface.co: