https://predibase.com/blog/teaching-ai-to-write-gpu-code-a-deep-dive-into-reinforcement-fine-tuning
#ai #nlp
#ai #nlp
Predibase
Train AI to Write GPU Code via Reinforcement Fine-Tuning
In this post, we’ll discuss how we taught an AI model to convert PyTorch code into efficient Triton kernels using a reinforcement learning algorithm inspired by PPO called Group Relative Preference Optimization (GRPO).
doing something
https://huggingface.co/collections/jbochi/madlad-400-65491e6a78726cac9a4b84b7 #nlp
Tested this neural network.
Better than OPUS and NLLB for Ukrainian.
3B model requires about 12 GB of memory (float32).
Better than OPUS and NLLB for Ukrainian.
3B model requires about 12 GB of memory (float32).
doing something
Tested this neural network. Better than OPUS and NLLB for Ukrainian. 3B model requires about 12 GB of memory (float32).
Sometimes we need hacks to fix NN's issues
https://github.com/pemistahl/lingua-py helps in such cases
https://github.com/pemistahl/lingua-py helps in such cases
As always, it's #daily_pain
Converting https://github.com/egorsmkv/cv10-uk-testset-clean/tree/main to HF dataset using sphn
#ai #speech
Converting https://github.com/egorsmkv/cv10-uk-testset-clean/tree/main to HF dataset using sphn
#ai #speech
doing something
As always, it's #daily_pain Converting https://github.com/egorsmkv/cv10-uk-testset-clean/tree/main to HF dataset using sphn #ai #speech
OK, a bug in the data preparation
Should be
Should be
'array': data[0], instead of 'array': data,Published the testset for Ukrainian to HF:
https://huggingface.co/datasets/Yehor/cv10-uk-testset-clean
#ai #speech
https://huggingface.co/datasets/Yehor/cv10-uk-testset-clean
#ai #speech
🔥2
doing something
Published the testset for Ukrainian to HF: https://huggingface.co/datasets/Yehor/cv10-uk-testset-clean #ai #speech
Here is an integration of inference code with https://huggingface.co/Yehor/w2v-bert-uk
Colab: https://drive.google.com/file/d/1vBGPGQsLsKI9Yy_MNeageyjK6x-cLehC/view?usp=sharing
#ai #asr #speech
Colab: https://drive.google.com/file/d/1vBGPGQsLsKI9Yy_MNeageyjK6x-cLehC/view?usp=sharing
#ai #asr #speech
https://huggingface.co/Yehor/w2v-xls-r-uk requires about 1.2 GB of GPU memory with float16.
It gives following metrics (without an external LM):
Accuracy on words: 79.76%
Accuracy on chars: 96.36%
Around 300 million of parameters.
#asr #ai #speech
It gives following metrics (without an external LM):
Accuracy on words: 79.76%
Accuracy on chars: 96.36%
Around 300 million of parameters.
#asr #ai #speech
❤1🔥1
doing something
Published the testset for Ukrainian to HF: https://huggingface.co/datasets/Yehor/cv10-uk-testset-clean #ai #speech