doing something
bs=24, whisper large-v3 https://github.com/egorsmkv/optimized-whisper #ai #asr
Tested a new Turbo model with HQQ, below an example how a ~2 hour audio is recognized
doing something
arestovich_ukrainian.txt
turbo_arestovich_ukrainian.txt
159.7 KB
- All Duration: 6465.6441
- All RTF: 0.003
- All elapsed: 19.2815
- All RTF: 0.003
- All elapsed: 19.2815
doing something
turbo_arestovich_ukrainian.txt
v3-arestovich.txt
153.1 KB
This test was on a 3090 card, so this file is also tested on the same hardware and software
RTF:
- V3: 0.0103
- V3 turbo: 0.003
RTF:
- V3: 0.0103
- V3 turbo: 0.003
https://colab.research.google.com/drive/1o9b2JQ8l9a39uOZZi9DWXQ15BXlCIfEu?usp=sharing
600M model with 4-bits uses about 1GB of VRAM
#nlp #ai #hqq
600M model with 4-bits uses about 1GB of VRAM
#nlp #ai #hqq
Google
Quantized NLLB using HQQ.ipynb
Colab notebook
doing something
decord can lag (it does not get the batch of frames) on GPU device if they have high values
This story has ended with a script that uses StreamReader from torchaudio.
I'm using the seek method to skip unrelated frames.
One important thing is that you need compiled ffmpeg with CUDA to do appropriate decoding.
I'm using the seek method to skip unrelated frames.
One important thing is that you need compiled ffmpeg with CUDA to do appropriate decoding.
An idea for a weekend project:
Adapt https://github.com/Gadersd/whisper-burn to the recently publish Whisper Turbo model.
Need to make only two things:
- update the repo to latest burn 🔥 version
- get down the number of decoder layers
#ai #asr
Adapt https://github.com/Gadersd/whisper-burn to the recently publish Whisper Turbo model.
Need to make only two things:
- update the repo to latest burn 🔥 version
- get down the number of decoder layers
#ai #asr
GitHub
GitHub - Gadersd/whisper-burn: A Rust implementation of OpenAI's Whisper model using the burn framework
A Rust implementation of OpenAI's Whisper model using the burn framework - Gadersd/whisper-burn
https://youtube.com/playlist?list=PLoROMvodv4rPOWA-omMM6STXaWW4FvJT8&feature=shared
https://deepgenerativemodels.github.io/
#ai #genai
https://deepgenerativemodels.github.io/
#ai #genai
YouTube
Stanford CS236: Deep Generative Models I 2023 I Stefano Ermon
For more information about Stanford's Artificial Intelligence programs visit: https://stanford.io/ai View the course website: https://deepgenerativemodels.gi...
Finally we can cut texts from memes to make other ones
https://huggingface.co/spaces/OzzyGT/diffusers-image-fill
#ai #genai
https://huggingface.co/spaces/OzzyGT/diffusers-image-fill
#ai #genai