This repository hosts the
IndicSeamless
model which is a SeamlessM4T-v2 finetuned on the
BhasaAnuvaad
dataset for
speech-to-text translation (STT)
across
Indian languages
. The dataset was filtered using the following thresholds before training:
Alignment Score
: 0.8
Mining Score
: 0.6
Performance Highlights
The model
outperforms the base SeamlessM4Tv2 model
and all competing STT systems, including cascaded approaches.
It
achieves a new SOTA on Fleurs and significantly surpasses all other systems on the BhasaAnuvaad test set, which includes a diverse range of data from new domains.
Model Usage
Installation
Ensure you have the required dependencies installed:
from datasets import load_dataset
import torch
defprocess_batch(batch):
audio_arrays = [audio["array"] for audio in batch["audio"]]
audio_inputs = processor(audio_arrays, sampling_rate=16_000, return_tensors="pt", padding=True).to("cuda")
text_outs = model.generate(**audio_inputs, tgt_lang="hin")
batch["predicted_text"] = [tokenizer.decode(text_out.cpu().numpy().squeeze(), clean_up_tokenization_spaces=True, skip_special_tokens=True) for text_out in text_outs]
return batch
defbatch_translate(language_code="hi_in", tgt_lang="hin"):
dataset = load_dataset("google/fleurs", language_code, split="test")
dataset = dataset.map(process_batch, batched=True, batch_size=8)
return dataset["predicted_text"]
# Example usage
target_language = "hi_in"
translations = batch_translate(target_language, tgt_lang="hin")
print(translations)
Citation
If you use BhasaAnuvaad in your work, please cite us:
@misc{jain2024bhasaanuvaadspeechtranslationdataset,
title={BhasaAnuvaad: A Speech Translation Dataset for 13 Indian Languages},
author={Sparsh Jain and Ashwin Sankar and Devilal Choudhary and Dhairya Suman and Nikhil Narasimhan and Mohammed Safi Ur Rahman Khan and Anoop Kunchukuttan and Mitesh M Khapra and Raj Dabre},
year={2024},
eprint={2411.04699},
archivePrefix={arXiv},
primaryClass={cs.CL},
url={https://arxiv.org/abs/2411.04699},
}
License
This model is released under the Creative Commons Attribution-NonCommercial 4.0 International (CC BY-NC 4.0) license.
Runs of ai4bharat indic-seamless on huggingface.co
3.8K
Total runs
0
24-hour runs
4
3-day runs
216
7-day runs
352
30-day runs
More Information About indic-seamless huggingface.co Model
indic-seamless huggingface.co is an AI model on huggingface.co that provides indic-seamless's model effect (), which can be used instantly with this ai4bharat indic-seamless model. huggingface.co supports a free trial of the indic-seamless model, and also provides paid use of the indic-seamless. Support call indic-seamless model through api, including Node.js, Python, http.
indic-seamless huggingface.co is an online trial and call api platform, which integrates indic-seamless's modeling effects, including api services, and provides a free online trial of indic-seamless, you can try indic-seamless online for free by clicking the link below.
ai4bharat indic-seamless online free url in huggingface.co:
indic-seamless is an open source model from GitHub that offers a free installation service, and any user can find indic-seamless on GitHub to install. At the same time, huggingface.co provides the effect of indic-seamless install, users can directly use indic-seamless installed effect in huggingface.co for debugging and trial. It also supports api for free installation.