cahya / wav2vec2-large-xlsr-breton

huggingface.co
Total runs: 26
24-hour runs: 0
7-day runs: 6
30-day runs: 5
Model's Last Updated: July 06 2021
automatic-speech-recognition

Introduction of wav2vec2-large-xlsr-breton

Model Details of wav2vec2-large-xlsr-breton

Wav2Vec2-Large-XLSR-Breton

Fine-tuned facebook/wav2vec2-large-xlsr-53 on the Breton Common Voice dataset . When using this model, make sure that your speech input is sampled at 16kHz.

Usage

The model can be used directly (without a language model) as follows:

import torch
import torchaudio
from datasets import load_dataset
from transformers import Wav2Vec2ForCTC, Wav2Vec2Processor
import re

test_dataset = load_dataset("common_voice", "br", split="test[:2%]")

processor = Wav2Vec2Processor.from_pretrained("cahya/wav2vec2-large-xlsr-breton")
model = Wav2Vec2ForCTC.from_pretrained("cahya/wav2vec2-large-xlsr-breton")

chars_to_ignore_regex = '[\\,\,\?\.\!\;\:\"\“\%\”\�\(\)\/\«\»\½\…]'

# Preprocessing the datasets.
# We need to read the aduio files as arrays
def speech_file_to_array_fn(batch):
    batch["sentence"] = re.sub(chars_to_ignore_regex, '', batch["sentence"]).lower() + " "
    batch["sentence"] = batch["sentence"].replace("ʼ", "'")
    batch["sentence"] = batch["sentence"].replace("’", "'")
    batch["sentence"] = batch["sentence"].replace('‘', "'")
    speech_array, sampling_rate = torchaudio.load(batch["path"])
    resampler = torchaudio.transforms.Resample(sampling_rate, 16_000)
    batch["speech"] = resampler(speech_array).squeeze().numpy()
    return batch

test_dataset = test_dataset.map(speech_file_to_array_fn)
inputs = processor(test_dataset[:2]["speech"], sampling_rate=16_000, return_tensors="pt", padding=True)

with torch.no_grad():
    logits = model(inputs.input_values, attention_mask=inputs.attention_mask).logits

predicted_ids = torch.argmax(logits, dim=-1)

print("Prediction:", processor.batch_decode(predicted_ids))
print("Reference:", test_dataset[:2]["sentence"])

The above code leads to the following prediction for the first two samples:

Prediction: ["ne' ler ket don a-benn us netra pa vez zer nic'hed evel-si", 'an eil hag egile']
Reference: ['"n\'haller ket dont a-benn eus netra pa vezer nec\'het evel-se." ', 'an eil hag egile. ']
Evaluation

The model can be evaluated as follows on the Breton test data of Common Voice.

import torch
import torchaudio
from datasets import load_dataset, load_metric
from transformers import Wav2Vec2ForCTC, Wav2Vec2Processor
import re

test_dataset = load_dataset("common_voice", "br", split="test")
wer = load_metric("wer")

processor = Wav2Vec2Processor.from_pretrained("cahya/wav2vec2-large-xlsr-breton")
model = Wav2Vec2ForCTC.from_pretrained("cahya/wav2vec2-large-xlsr-breton") 
model.to("cuda")

chars_to_ignore_regex = '[\\,\,\?\.\!\;\:\"\“\%\”\�\(\)\/\«\»\½\…]'

# Preprocessing the datasets.
# We need to read the aduio files as arrays
def speech_file_to_array_fn(batch):
    batch["sentence"] = re.sub(chars_to_ignore_regex, '', batch["sentence"]).lower() + " "
    batch["sentence"] = batch["sentence"].replace("ʼ", "'")
    batch["sentence"] = batch["sentence"].replace("’", "'")
    batch["sentence"] = batch["sentence"].replace('‘', "'")
    speech_array, sampling_rate = torchaudio.load(batch["path"])
    resampler = torchaudio.transforms.Resample(sampling_rate, 16_000)
    batch["speech"] = resampler(speech_array).squeeze().numpy()
    return batch

test_dataset = test_dataset.map(speech_file_to_array_fn)

# Preprocessing the datasets.
# We need to read the aduio files as arrays
def evaluate(batch):
    inputs = processor(batch["speech"], sampling_rate=16_000, return_tensors="pt", padding=True)

    with torch.no_grad():
        logits = model(inputs.input_values.to("cuda"), attention_mask=inputs.attention_mask.to("cuda")).logits

    pred_ids = torch.argmax(logits, dim=-1)
    batch["pred_strings"] = processor.batch_decode(pred_ids)
    return batch

result = test_dataset.map(evaluate, batched=True, batch_size=8)

print("WER: {:2f}".format(100 * wer.compute(predictions=result["pred_strings"], references=result["sentence"])))

Test Result : 41.71 %

Training

The Common Voice train , validation , and ... datasets were used for training as well as ... and ... # TODO

The script used for training can be found here (will be available soon)

Runs of cahya wav2vec2-large-xlsr-breton on huggingface.co

26
Total runs
0
24-hour runs
3
3-day runs
6
7-day runs
5
30-day runs

More Information About wav2vec2-large-xlsr-breton huggingface.co Model

More wav2vec2-large-xlsr-breton license Visit here:

https://choosealicense.com/licenses/apache-2.0

wav2vec2-large-xlsr-breton huggingface.co

wav2vec2-large-xlsr-breton huggingface.co is an AI model on huggingface.co that provides wav2vec2-large-xlsr-breton's model effect (), which can be used instantly with this cahya wav2vec2-large-xlsr-breton model. huggingface.co supports a free trial of the wav2vec2-large-xlsr-breton model, and also provides paid use of the wav2vec2-large-xlsr-breton. Support call wav2vec2-large-xlsr-breton model through api, including Node.js, Python, http.

wav2vec2-large-xlsr-breton huggingface.co Url

https://huggingface.co/cahya/wav2vec2-large-xlsr-breton

cahya wav2vec2-large-xlsr-breton online free

wav2vec2-large-xlsr-breton huggingface.co is an online trial and call api platform, which integrates wav2vec2-large-xlsr-breton's modeling effects, including api services, and provides a free online trial of wav2vec2-large-xlsr-breton, you can try wav2vec2-large-xlsr-breton online for free by clicking the link below.

cahya wav2vec2-large-xlsr-breton online free url in huggingface.co:

https://huggingface.co/cahya/wav2vec2-large-xlsr-breton

wav2vec2-large-xlsr-breton install

wav2vec2-large-xlsr-breton is an open source model from GitHub that offers a free installation service, and any user can find wav2vec2-large-xlsr-breton on GitHub to install. At the same time, huggingface.co provides the effect of wav2vec2-large-xlsr-breton install, users can directly use wav2vec2-large-xlsr-breton installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

wav2vec2-large-xlsr-breton install url in huggingface.co:

https://huggingface.co/cahya/wav2vec2-large-xlsr-breton

Url of wav2vec2-large-xlsr-breton

wav2vec2-large-xlsr-breton huggingface.co Url

Provider of wav2vec2-large-xlsr-breton huggingface.co

cahya
ORGANIZATIONS

Other API from cahya

huggingface.co

Total runs: 85
Run Growth: -44
Growth Rate: -53.66%
Updated:March 01 2025
huggingface.co

Total runs: 74
Run Growth: 61
Growth Rate: 82.43%
Updated:March 24 2022
huggingface.co

Total runs: 24
Run Growth: 18
Growth Rate: 75.00%
Updated:February 01 2022
huggingface.co

Total runs: 15
Run Growth: 9
Growth Rate: 64.29%
Updated:June 24 2024
huggingface.co

Total runs: 12
Run Growth: -5
Growth Rate: -41.67%
Updated:February 19 2023
huggingface.co

Total runs: 10
Run Growth: -132
Growth Rate: -1320.00%
Updated:April 17 2025
huggingface.co

Total runs: 9
Run Growth: -23
Growth Rate: -255.56%
Updated:November 24 2022
huggingface.co

Total runs: 7
Run Growth: -3
Growth Rate: -42.86%
Updated:February 01 2023
huggingface.co

Total runs: 6
Run Growth: 1
Growth Rate: 16.67%
Updated:July 06 2021
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:February 04 2025
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:March 12 2023
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:May 21 2024
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:February 04 2025
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:February 11 2025