doing something
why it is so painful... https://huggingface.co/datasets/Helsinki-NLP/opus-100 #nlp
GitHub
GitHub - MicrosoftTranslator/NTREX: NTREX -- News Test References for MT Evaluation
NTREX -- News Test References for MT Evaluation. Contribute to MicrosoftTranslator/NTREX development by creating an account on GitHub.
https://predibase.com/blog/teaching-ai-to-write-gpu-code-a-deep-dive-into-reinforcement-fine-tuning
#ai #nlp
#ai #nlp
Predibase
Train AI to Write GPU Code via Reinforcement Fine-Tuning
In this post, we’ll discuss how we taught an AI model to convert PyTorch code into efficient Triton kernels using a reinforcement learning algorithm inspired by PPO called Group Relative Preference Optimization (GRPO).
doing something
https://huggingface.co/collections/jbochi/madlad-400-65491e6a78726cac9a4b84b7 #nlp
Tested this neural network.
Better than OPUS and NLLB for Ukrainian.
3B model requires about 12 GB of memory (float32).
Better than OPUS and NLLB for Ukrainian.
3B model requires about 12 GB of memory (float32).
doing something
Tested this neural network. Better than OPUS and NLLB for Ukrainian. 3B model requires about 12 GB of memory (float32).
Sometimes we need hacks to fix NN's issues
https://github.com/pemistahl/lingua-py helps in such cases
https://github.com/pemistahl/lingua-py helps in such cases
As always, it's #daily_pain
Converting https://github.com/egorsmkv/cv10-uk-testset-clean/tree/main to HF dataset using sphn
#ai #speech
Converting https://github.com/egorsmkv/cv10-uk-testset-clean/tree/main to HF dataset using sphn
#ai #speech
doing something
As always, it's #daily_pain Converting https://github.com/egorsmkv/cv10-uk-testset-clean/tree/main to HF dataset using sphn #ai #speech
OK, a bug in the data preparation
Should be
Should be
'array': data[0], instead of 'array': data,