BART-SLED (SLiding-Encoder and Decoder, base-sized model)
SLED models use pretrained, short-range encoder-decoder models, and apply them over
long-text inputs by splitting the input into multiple overlapping chunks, encoding each independently and perform fusion-in-decoder
Model description
This SLED model is based on the BART model, which is described in its
model card
.
BART is particularly effective when fine-tuned for text generation (e.g. summarization, translation) but also works
well for comprehension tasks (e.g. text classification, question answering). When used as a BART-SLED model, it can be applied on long text tasks.
Once installed, SLED is fully compatible with HuggingFace's AutoClasses (AutoTokenizer, AutoConfig, AutoModel
and AutoModelForCausalLM) and can be loaded using the from_pretrained methods
import sled # *** required so that SledModels will be registered for the AutoClasses ***
model = AutoModel.from_pretrained('tau/bart-base-sled')
Here is how to use this model in PyTorch:
from sled import SledTokenizer, SledModel
tokenizer = SledTokenizer.from_pretrained('tau/bart-base-sled')
model = SledModel.from_pretrained('tau/bart-base-sled')
inputs = tokenizer("Hello, my dog is cute", return_tensors="pt")
outputs = model(**inputs)
last_hidden_states = outputs.last_hidden_state
You can also replace SledModel by SledModelForConditionalGeneration for Seq2Seq generation
model = SledModelForConditionalGeneration.from_pretrained('tau/bart-base-sled')
In case you wish to apply SLED on a task containing a prefix (e.g. question) which should be given as a context to
every chunk, you can pass the
prefix_length
tensor input as well (A LongTensor in the length of the batch size).
import torch
import sled # *** required so that SledModels will be registered for the AutoClasses ***
tokenizer = AutoTokenizer.from_pretrained('tau/bart-base-sled')
model = AutoModel.from_pretrained('tau/bart-base-sled')
document_input_ids = tokenizer("Dogs are great for you.", return_tensors="pt").input_ids
prefix_input_ids = tokenizer("Are dogs good for you?", return_tensors="pt").input_ids
input_ids = torch.cat((prefix_input_ids, document_input_ids), dim=-1)
attention_mask = torch.ones_like(input_ids)
prefix_length = torch.LongTensor([[prefix_input_ids.size(1)]])
outputs = model(input_ids=input_ids, attention_mask=attention_mask, prefix_length=prefix_length)
last_hidden_states = outputs.last_hidden_state
BibTeX entry and citation info
Please cite both the SLED
paper
and the BART
paper
by Lewis et al as well as GovReport by Huang et al
@inproceedings{Ivgi2022EfficientLU,
title={Efficient Long-Text Understanding with Short-Text Models},
author={Maor Ivgi and Uri Shaham and Jonathan Berant},
year={2022}
}
@article{DBLP:journals/corr/abs-1910-13461,
author = {Mike Lewis and
Yinhan Liu and
Naman Goyal and
Marjan Ghazvininejad and
Abdelrahman Mohamed and
Omer Levy and
Veselin Stoyanov and
Luke Zettlemoyer},
title = {{BART:} Denoising Sequence-to-Sequence Pre-training for Natural Language
Generation, Translation, and Comprehension},
journal = {CoRR},
volume = {abs/1910.13461},
year = {2019},
url = {http://arxiv.org/abs/1910.13461},
eprinttype = {arXiv},
eprint = {1910.13461},
timestamp = {Thu, 31 Oct 2019 14:02:26 +0100},
biburl = {https://dblp.org/rec/journals/corr/abs-1910-13461.bib},
bibsource = {dblp computer science bibliography, https://dblp.org}
}
@inproceedings{huang2021govreport,
title = "Efficient Attentions for Long Document Summarization",
author = "Huang, Luyang and
Cao, Shuyang and
Parulian, Nikolaus and
Ji, Heng and
Wang, Lu",
booktitle = "Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies",
month = jun,
year = "2021",
address = "Online",
publisher = "Association for Computational Linguistics",
url = "https://aclanthology.org/2021.naacl-main.112",
doi = "10.18653/v1/2021.naacl-main.112",
pages = "1419--1436"
}
Runs of tau bart-base-sled-govreport on huggingface.co
43
Total runs
0
24-hour runs
0
3-day runs
5
7-day runs
34
30-day runs
More Information About bart-base-sled-govreport huggingface.co Model
bart-base-sled-govreport huggingface.co is an AI model on huggingface.co that provides bart-base-sled-govreport's model effect (), which can be used instantly with this tau bart-base-sled-govreport model. huggingface.co supports a free trial of the bart-base-sled-govreport model, and also provides paid use of the bart-base-sled-govreport. Support call bart-base-sled-govreport model through api, including Node.js, Python, http.
bart-base-sled-govreport huggingface.co is an online trial and call api platform, which integrates bart-base-sled-govreport's modeling effects, including api services, and provides a free online trial of bart-base-sled-govreport, you can try bart-base-sled-govreport online for free by clicking the link below.
tau bart-base-sled-govreport online free url in huggingface.co:
bart-base-sled-govreport is an open source model from GitHub that offers a free installation service, and any user can find bart-base-sled-govreport on GitHub to install. At the same time, huggingface.co provides the effect of bart-base-sled-govreport install, users can directly use bart-base-sled-govreport installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
bart-base-sled-govreport install url in huggingface.co: