doing something
710 subscribers
908 photos
36 videos
36 files
1.94K links
@smlkw doing something, just my notes to keep updates
Download Telegram
Audio
Japanese version with cloned voice
doing something
Gemini helps to visualize
Then helps to understand pronunciation for lip motion
A way to convert datasets from JSONL to DuckDB quickly.

#duckdb
doing something
A way to convert datasets from JSONL to DuckDB quickly. #duckdb
That example shows this dataset: https://huggingface.co/datasets/speech-uk/generated-news-comments

The same method was used to create a table for https://huggingface.co/datasets/speech-uk/voiced-news-comments dataset

DuckDB gives ability to fastly prototype

In the Speech-UK initiative these two datasets will be used to crowd-source data for Ukrainian ASR

Next step: create a Telegram bot using https://github.com/AmarnathCJD/gogram to make voice crowd-sourcing available to beta users
Hmm…
When I implement tasks in my Speech-UK initiative (btw, follow it here - https://huggingface.co/speech-uk) some projects arise like this one.

It gives you ability to filter out bad samples from a dataset based on another model Meta made recently.

https://github.com/egorsmkv/audiobox-aesthetics-inference

#ml #audio
NSA at home:
😁1🌚1
doing something
https://github.com/facebook/sapling #rust
They don't have a prebuilt package for Ubuntu 24, so I am building it from the source code

It has a lot of dependencies, btw