doing something
Then helps to understand pronunciation for lip motion
This media is not supported in your browser
VIEW IN TELEGRAM
Some preprocessing in Audacity and CLI magic with a video note and we have a fake
doing something
A way to convert datasets from JSONL to DuckDB quickly. #duckdb
That example shows this dataset: https://huggingface.co/datasets/speech-uk/generated-news-comments
The same method was used to create a table for https://huggingface.co/datasets/speech-uk/voiced-news-comments dataset
DuckDB gives ability to fastly prototype
In the Speech-UK initiative these two datasets will be used to crowd-source data for Ukrainian ASR
Next step: create a Telegram bot using https://github.com/AmarnathCJD/gogram to make voice crowd-sourcing available to beta users
The same method was used to create a table for https://huggingface.co/datasets/speech-uk/voiced-news-comments dataset
DuckDB gives ability to fastly prototype
In the Speech-UK initiative these two datasets will be used to crowd-source data for Ukrainian ASR
Next step: create a Telegram bot using https://github.com/AmarnathCJD/gogram to make voice crowd-sourcing available to beta users
When I implement tasks in my Speech-UK initiative (btw, follow it here - https://huggingface.co/speech-uk) some projects arise like this one.
It gives you ability to filter out bad samples from a dataset based on another model Meta made recently.
https://github.com/egorsmkv/audiobox-aesthetics-inference
#ml #audio
It gives you ability to filter out bad samples from a dataset based on another model Meta made recently.
https://github.com/egorsmkv/audiobox-aesthetics-inference
#ml #audio
huggingface.co
speech-uk (Speech-UK initiative)
We are driving innovation in Ukrainian speech technology 🇺🇦
doing something
When I implement tasks in my Speech-UK initiative (btw, follow it here - https://huggingface.co/speech-uk) some projects arise like this one. It gives you ability to filter out bad samples from a dataset based on another model Meta made recently. https…
Also, added a Gradio app to https://github.com/egorsmkv/marblenet-inference
You can perform VAD operation on large files (if you have a GPU) with ease now
#audio
You can perform VAD operation on large files (if you have a GPU) with ease now
#audio
GitHub
GitHub - egorsmkv/marblenet-inference: Inference code for Frame MarbleNet (VAD from NeMo)
Inference code for Frame MarbleNet (VAD from NeMo) - egorsmkv/marblenet-inference
doing something
When I implement tasks in my Speech-UK initiative (btw, follow it here - https://huggingface.co/speech-uk) some projects arise like this one. It gives you ability to filter out bad samples from a dataset based on another model Meta made recently. https…
Added stats over metrics to better understand distributions
doing something
https://github.com/facebook/sapling #rust
They don't have a prebuilt package for Ubuntu 24, so I am building it from the source code
It has a lot of dependencies, btw
It has a lot of dependencies, btw