ai4bharat / indic-seamless

huggingface.co
Total runs: 3.8K
24-hour runs: 0
7-day runs: 216
30-day runs: 352
Model's Last Updated: March 07 2025
automatic-speech-recognition

Introduction of indic-seamless

Model Details of indic-seamless

IndicSeamless for Speech-to-Text Translation

Open in HuggingFace
Model Overview

This repository hosts the IndicSeamless model which is a SeamlessM4T-v2 finetuned on the BhasaAnuvaad dataset for speech-to-text translation (STT) across Indian languages . The dataset was filtered using the following thresholds before training:

  • Alignment Score : 0.8
  • Mining Score : 0.6
Performance Highlights
  • The model outperforms the base SeamlessM4Tv2 model and all competing STT systems, including cascaded approaches.
  • It achieves a new SOTA on Fleurs and significantly surpasses all other systems on the BhasaAnuvaad test set, which includes a diverse range of data from new domains.
Model Usage
Installation

Ensure you have the required dependencies installed:

pip install torch torchaudio transformers datasets
Loading the Model
import torchaudio
from transformers import SeamlessM4Tv2ForSpeechToText
from transformers import SeamlessM4TTokenizer, SeamlessM4TFeatureExtractor

model = SeamlessM4Tv2ForSpeechToText.from_pretrained("ai4bharat/indic-seamless").to("cuda")
processor = SeamlessM4TFeatureExtractor.from_pretrained("ai4bharat/indic-seamless")
tokenizer = SeamlessM4TTokenizer.from_pretrained("ai4bharat/indic-seamless")
Single Audio Inference
audio, orig_freq = torchaudio.load("../10002398547238927970.wav")
audio = torchaudio.functional.resample(audio, orig_freq=orig_freq, new_freq=16_000) # must be a 16 kHz waveform array
audio_inputs = processor(audio, sampling_rate=16_000, return_tensors="pt").to("cuda")

text_out = model.generate(**audio_inputs, tgt_lang="hin")[0].cpu().numpy().squeeze()
print(tokenizer.decode(text_out, clean_up_tokenization_spaces=True, skip_special_tokens=True))
Inference on Fleurs Dataset
from datasets import load_dataset

dataset = load_dataset("google/fleurs", "hi_in", split="test")

def process_audio(example):
    audio = example["audio"]["array"]
    audio_inputs = processor(audio, sampling_rate=16_000, return_tensors="pt").to("cuda")
    text_out = model.generate(**audio_inputs, tgt_lang="hin")[0].cpu().numpy().squeeze()
    return {"predicted_text": tokenizer.decode(text_out, clean_up_tokenization_spaces=True, skip_special_tokens=True)}

dataset = dataset.map(process_audio)
dataset = dataset.remove_columns(["audio"])
dataset.to_csv("fleurs_hi_predictions.csv")
Batch Translation using Fleurs
from datasets import load_dataset
import torch

def process_batch(batch):
    audio_arrays = [audio["array"] for audio in batch["audio"]]
    audio_inputs = processor(audio_arrays, sampling_rate=16_000, return_tensors="pt", padding=True).to("cuda")
    text_outs = model.generate(**audio_inputs, tgt_lang="hin")
    batch["predicted_text"] = [tokenizer.decode(text_out.cpu().numpy().squeeze(), clean_up_tokenization_spaces=True, skip_special_tokens=True) for text_out in text_outs]
    return batch

def batch_translate(language_code="hi_in", tgt_lang="hin"):
    dataset = load_dataset("google/fleurs", language_code, split="test")
    dataset = dataset.map(process_batch, batched=True, batch_size=8)
    return dataset["predicted_text"]

# Example usage
target_language = "hi_in"
translations = batch_translate(target_language, tgt_lang="hin")
print(translations)
Citation

If you use BhasaAnuvaad in your work, please cite us:

@misc{jain2024bhasaanuvaadspeechtranslationdataset,
      title={BhasaAnuvaad: A Speech Translation Dataset for 13 Indian Languages}, 
      author={Sparsh Jain and Ashwin Sankar and Devilal Choudhary and Dhairya Suman and Nikhil Narasimhan and Mohammed Safi Ur Rahman Khan and Anoop Kunchukuttan and Mitesh M Khapra and Raj Dabre},
      year={2024},
      eprint={2411.04699},
      archivePrefix={arXiv},
      primaryClass={cs.CL},
      url={https://arxiv.org/abs/2411.04699}, 
}
License

This model is released under the Creative Commons Attribution-NonCommercial 4.0 International (CC BY-NC 4.0) license.

Runs of ai4bharat indic-seamless on huggingface.co

3.8K
Total runs
0
24-hour runs
4
3-day runs
216
7-day runs
352
30-day runs

More Information About indic-seamless huggingface.co Model

More indic-seamless license Visit here:

https://choosealicense.com/licenses/cc-by-nc-4.0

indic-seamless huggingface.co

indic-seamless huggingface.co is an AI model on huggingface.co that provides indic-seamless's model effect (), which can be used instantly with this ai4bharat indic-seamless model. huggingface.co supports a free trial of the indic-seamless model, and also provides paid use of the indic-seamless. Support call indic-seamless model through api, including Node.js, Python, http.

indic-seamless huggingface.co Url

https://huggingface.co/ai4bharat/indic-seamless

ai4bharat indic-seamless online free

indic-seamless huggingface.co is an online trial and call api platform, which integrates indic-seamless's modeling effects, including api services, and provides a free online trial of indic-seamless, you can try indic-seamless online for free by clicking the link below.

ai4bharat indic-seamless online free url in huggingface.co:

https://huggingface.co/ai4bharat/indic-seamless

indic-seamless install

indic-seamless is an open source model from GitHub that offers a free installation service, and any user can find indic-seamless on GitHub to install. At the same time, huggingface.co provides the effect of indic-seamless install, users can directly use indic-seamless installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

indic-seamless install url in huggingface.co:

https://huggingface.co/ai4bharat/indic-seamless

Url of indic-seamless

indic-seamless huggingface.co Url

Provider of indic-seamless huggingface.co

ai4bharat
ORGANIZATIONS

Other API from ai4bharat

huggingface.co

Total runs: 245.2K
Run Growth: -3.6K
Growth Rate: -1.46%
Updated:August 08 2022
huggingface.co

Total runs: 31.8K
Run Growth: 5.0K
Growth Rate: 15.82%
Updated:March 03 2026
huggingface.co

Total runs: 10.0K
Run Growth: 1.1K
Growth Rate: 10.93%
Updated:December 21 2022
huggingface.co

Total runs: 3.6K
Run Growth: 3.1K
Growth Rate: 85.69%
Updated:November 19 2025
huggingface.co

Total runs: 3.1K
Run Growth: -9.8K
Growth Rate: -317.34%
Updated:August 08 2022
huggingface.co

Total runs: 1.4K
Run Growth: -86
Growth Rate: -5.46%
Updated:March 11 2024
huggingface.co

Total runs: 316
Run Growth: 214
Growth Rate: 69.03%
Updated:June 01 2022
huggingface.co

Total runs: 109
Run Growth: 65
Growth Rate: 53.72%
Updated:September 01 2026
huggingface.co

Total runs: 92
Run Growth: 29
Growth Rate: 31.87%
Updated:October 18 2024
huggingface.co

Total runs: 90
Run Growth: 28
Growth Rate: 31.46%
Updated:October 18 2024
huggingface.co

Total runs: 85
Run Growth: 32
Growth Rate: 38.55%
Updated:October 18 2024
huggingface.co

Total runs: 85
Run Growth: 28
Growth Rate: 34.57%
Updated:October 18 2024
huggingface.co

Total runs: 84
Run Growth: 22
Growth Rate: 27.16%
Updated:October 18 2024
huggingface.co

Total runs: 82
Run Growth: 19
Growth Rate: 23.46%
Updated:October 18 2024