doing something
Published Open Source Crimean Tatar Text-to-Speech dataset to Hugging Face: https://huggingface.co/datasets/Yehor/qirimtatar-tts #ai #speech
Now, it's in different datsets to make it more useful (with train/test splits):
- https://huggingface.co/datasets/speech-uk/qirimtatar-abibullah-tts
- https://huggingface.co/datasets/speech-uk/qirimtatar-sevil-tts
- https://huggingface.co/datasets/speech-uk/qirimtatar-arslan-tts
#ai #speech
- https://huggingface.co/datasets/speech-uk/qirimtatar-abibullah-tts
- https://huggingface.co/datasets/speech-uk/qirimtatar-sevil-tts
- https://huggingface.co/datasets/speech-uk/qirimtatar-arslan-tts
#ai #speech
huggingface.co
speech-uk/tts-crh-abibullah Β· Datasets at Hugging Face
Weβre on a journey to advance and democratize artificial intelligence through open source and open science.
π1
Ukrainian voices for Text-to-Speech now in separated datasets:
https://huggingface.co/collections/speech-uk/ukrainian-text-to-speech-67bd059d61b2598f3a2a7969
#ai #speech
https://huggingface.co/collections/speech-uk/ukrainian-text-to-speech-67bd059d61b2598f3a2a7969
#ai #speech
huggingface.co
Ukrainian Text-to-Speech datasets - a speech-uk Collection
Five voices: Mykyta, Oleksa, Lada, Kateryna or Tetiana
π₯2
Added a section with phonemizers to https://github.com/egorsmkv/speech-recognition-uk
https://github.com/egorsmkv/speech-recognition-uk/issues/57
#ai #speech
https://github.com/egorsmkv/speech-recognition-uk/issues/57
#ai #speech
GitHub
GitHub - egorsmkv/speech-recognition-uk: πΊπ¦ Speech Recognition & Synthesis for Ukrainian
πΊπ¦ Speech Recognition & Synthesis for Ukrainian. Contribute to egorsmkv/speech-recognition-uk development by creating an account on GitHub.
Made a space to simplify my ASR research:
https://huggingface.co/spaces/Yehor/evaluate-asr-outputs
#ai #speech
https://huggingface.co/spaces/Yehor/evaluate-asr-outputs
#ai #speech
π1π₯1
https://huggingface.co/spaces/fastrtc/object-detection
https://huggingface.co/spaces/fastrtc/object-detection/blob/main/app.py
#cv #ai
https://huggingface.co/spaces/fastrtc/object-detection/blob/main/app.py
#cv #ai
huggingface.co
Object Detection - a Hugging Face Space by fastrtc
This app detects objects in real-time video streams. Users provide a live video feed, and the app highlights detected objects with confidence scores.
doing something
Batch inference using vllm #ai #vllm #speech
Added batch mode calculation to https://huggingface.co/spaces/Yehor/evaluate-asr-outputs
Btw, batched vllm inference gives better RTF for Whisper
#asr #ai #speech
Btw, batched vllm inference gives better RTF for Whisper
#asr #ai #speech
vllm + A100 (40 GB in Google Colab) + https://huggingface.co/Yehor/whisper-large-v3-turbo-quantized-uk using batch_size=128
#ai #speech #whisper
#ai #speech #whisper
π₯2
Initial work to add OWSM models from Espnet:
https://github.com/egorsmkv/speech-recognition-uk/issues/69
#ai #speech
https://github.com/egorsmkv/speech-recognition-uk/issues/69
#ai #speech
GitHub
Add OWSM-CTC models Β· Issue #69 Β· egorsmkv/speech-recognition-uk
https://huggingface.co/collections/espnet/owsm-ctc-ultra-fast-speech-foundation-models-67ab7a9d91a82bab571ca42c