tau / t5-v1_1-large-rss

huggingface.co
Total runs: 8
24-hour runs: 0
7-day runs: 3
30-day runs: 6
Model's Last Updated: August 21 2021
text-generation

Introduction of t5-v1_1-large-rss

Model Details of t5-v1_1-large-rss

T5-V1.1-large-rss

This model is T5-v1.1-large finetuned on RSS dataset. The model was finetuned as part of "How Optimal is Greedy Decoding for Extractive Question Answering?" , while the RSS pretraining method was introduced in this paper .

Model description

The original T5-v1.1-large was only pre-trained on C4 excluding any supervised training. Our version is further trained on Rucurrent Span Selection scheme (RSS), using a sample from the dataset used to pretrain Splinter :

  • contexts with a span occurring more than once are detected
  • a single instance of the recurring span is maked
  • the model is trained (teacher forcing) to predict the masked span This training scheme naturally matches the extractive question answering task.

During training time, the masked span is replaced with <extra_id_0> and the labels are formatted as <extra_id_0>span<extra_id_0> . Unlike Splinter , only one span is mask at a time.

Intended uses & limitations

This model naturally fits tasks where a span from a context is intended to be copied, like extractive question answering. This checkpoint is primarily aimed to be used in zero-shot setting - further fine-tuning it on an annotated dataset gives equal results to those of the original T5-v1.1-large.

How to use

You can use this model directly but it is recommended to format the input to be aligned with that of the training scheme and as a text-question context:

from transformers import  AutoModelForSeq2SeqLM, AutoTokenizer
model = AutoModelForSeq2SeqLM.from_pretrained('tau/t5-v1_1-large-rss')
tokenizer = AutoTokenizer.from_pretrained('tau/t5-v1_1-large-rss')

passage = 'Barack Hussein Obama II is an American politician and attorney who served as the 44th president of the United States from 2009 to 2017. '
question = 'When was Obama inaugurated?'
text = f'Text: {passage}.\nQuestion: {question}\nAnswer:{tokenizer.additional_special_tokens[0]}.'
encoded_input = tokenizer(text, return_tensors='pt')
output_ids = model.generate(input_ids=encoded_input.input_ids, attention_mask=encoded_input.attention_mask,
               eos_token_id=tokenizer.additional_special_tokens_ids[1], num_beams=1, max_length=512, min_length=3)
tokenizer.decode(output_ids[0])

The generated answer is then "<pad><extra_id_0> 2009<extra_id_1>" , while the one generated by the original T5-v1.1-large is "<pad><extra_id_0> On January 20, 2009<extra_id_1>" - a correct yet non-extractive answer.

Limitations and bias

Although using the model with greedy decoding tends toward extracted outputs, is may sometimes produce non-extracted ones - may it be different casing or a whole different string (or substring) that may bear another semantic meaning.

Pretraining

The model was finetuned with 100,000 rss-examples for 3 epochs using Adafactor optimizer with constant learning rate of 5e-5.

Evaluation results

Evaluated over few-shot QA in a zero-shot setting (no finetuning on annotated examples):

Model \ Dataset SQuAD TriviaQA NaturalQs NewsQA SearchQA HotpotQA BioASQ TextbookQA
T5 50.4 61.7 42.1 19.2 24.0 43.3 55.5 17.8
T5-rss 71.4 69.3 57.2 43.2 29.7 59.0 65.5 39.0

The gap between the two models diminishes as more training examples are introduced, for additional result see the paper .

BibTeX entry and citation info
@inproceedings{ram-etal-2021-shot,
    title = "Few-Shot Question Answering by Pretraining Span Selection",
    author = "Ram, Ori  and
      Kirstain, Yuval  and
      Berant, Jonathan  and
      Globerson, Amir  and
      Levy, Omer",
    booktitle = "Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers)",
    month = aug,
    year = "2021",
    address = "Online",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2021.acl-long.239",
    doi = "10.18653/v1/2021.acl-long.239",
    pages = "3066--3079",
},
@misc{castel2021optimal,
      title={How Optimal is Greedy Decoding for Extractive Question Answering?}, 
      author={Or Castel and Ori Ram and Avia Efrat and Omer Levy},
      year={2021},
      eprint={2108.05857},
      archivePrefix={arXiv},
      primaryClass={cs.CL}
}

Runs of tau t5-v1_1-large-rss on huggingface.co

8
Total runs
0
24-hour runs
0
3-day runs
3
7-day runs
6
30-day runs

More Information About t5-v1_1-large-rss huggingface.co Model

t5-v1_1-large-rss huggingface.co

t5-v1_1-large-rss huggingface.co is an AI model on huggingface.co that provides t5-v1_1-large-rss's model effect (), which can be used instantly with this tau t5-v1_1-large-rss model. huggingface.co supports a free trial of the t5-v1_1-large-rss model, and also provides paid use of the t5-v1_1-large-rss. Support call t5-v1_1-large-rss model through api, including Node.js, Python, http.

t5-v1_1-large-rss huggingface.co Url

https://huggingface.co/tau/t5-v1_1-large-rss

tau t5-v1_1-large-rss online free

t5-v1_1-large-rss huggingface.co is an online trial and call api platform, which integrates t5-v1_1-large-rss's modeling effects, including api services, and provides a free online trial of t5-v1_1-large-rss, you can try t5-v1_1-large-rss online for free by clicking the link below.

tau t5-v1_1-large-rss online free url in huggingface.co:

https://huggingface.co/tau/t5-v1_1-large-rss

t5-v1_1-large-rss install

t5-v1_1-large-rss is an open source model from GitHub that offers a free installation service, and any user can find t5-v1_1-large-rss on GitHub to install. At the same time, huggingface.co provides the effect of t5-v1_1-large-rss install, users can directly use t5-v1_1-large-rss installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

t5-v1_1-large-rss install url in huggingface.co:

https://huggingface.co/tau/t5-v1_1-large-rss

Url of t5-v1_1-large-rss

t5-v1_1-large-rss huggingface.co Url

Provider of t5-v1_1-large-rss huggingface.co

tau
ORGANIZATIONS

Other API from tau

huggingface.co

Total runs: 28.8K
Run Growth: 27.0K
Growth Rate: 93.85%
Updated:August 17 2021
huggingface.co

Total runs: 548
Run Growth: 400
Growth Rate: 72.99%
Updated:August 10 2022
huggingface.co

Total runs: 60
Run Growth: 0
Growth Rate: 0.00%
Updated:February 09 2022
huggingface.co

Total runs: 17
Run Growth: 10
Growth Rate: 58.82%
Updated:August 17 2021
huggingface.co

Total runs: 6
Run Growth: -10
Growth Rate: -166.67%
Updated:August 10 2022
huggingface.co

Total runs: 5
Run Growth: 0
Growth Rate: 0.00%
Updated:April 09 2022
huggingface.co

Total runs: 4
Run Growth: 0
Growth Rate: 0.00%
Updated:March 07 2022
huggingface.co

Total runs: 3
Run Growth: 0
Growth Rate: 0.00%
Updated:April 09 2022
huggingface.co

Total runs: 3
Run Growth: 0
Growth Rate: 0.00%
Updated:March 14 2022
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:March 07 2022
huggingface.co

Total runs: 0
Run Growth: -7
Growth Rate: 0.00%
Updated:May 08 2022
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:October 21 2021