Automatic Speech Recognition
Transformers
Safetensors
seamless_m4t_v2
feature-extraction
audio-to-audio
text-to-speech
seamless_communication
Instructions to use facebook/seamless-m4t-v2-large with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use facebook/seamless-m4t-v2-large with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("automatic-speech-recognition", model="facebook/seamless-m4t-v2-large")# pip install -U transformers accelerate # Load model directly from transformers import AutoTokenizer, AutoModel tokenizer = AutoTokenizer.from_pretrained("facebook/seamless-m4t-v2-large") model = AutoModel.from_pretrained("facebook/seamless-m4t-v2-large", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Inference with finetuned model (.pt format) which is created using the finetuning in github
#49
by ThivyanRR - opened
Hi, I am new to this field. I want to finetune the v2_large model for some (non-supported) languages to run TTS. Where can I find the documentation for that? The steps mentioned in github creates a .pt file but where can i create the config.json file and others etc. and How can we infer with the finetuned model. Kindly help me in this as I am new to this domain.
Hello brother can you get any update regarding this