Fast denoiser based on diffusions:
https://github.com/sp-uhh/sgmse_crp
RTF about 0.0375 on RTX 4k Ada, because it needs only 1 reverse step
#ai #diffusions #speech
https://github.com/sp-uhh/sgmse_crp
RTF about 0.0375 on RTX 4k Ada, because it needs only 1 reverse step
#ai #diffusions #speech
GitHub
GitHub - sp-uhh/sgmse_crp
Contribute to sp-uhh/sgmse_crp development by creating an account on GitHub.
We’ve been talking about PDF processing yesterday at UDS and today I’ve discovered it:
https://x.com/hu_yifei/status/1828870309857915341?s=46&t=7jwH29MvU0R301CgvqVBYw
#cv #ai
https://x.com/hu_yifei/status/1828870309857915341?s=46&t=7jwH29MvU0R301CgvqVBYw
#cv #ai
This author has another model (based on Florence-2) that extracts parts in the article: https://huggingface.co/yifeihu/TFT-ID-1.0
In the album some examples how it works. Next we can OCR these images and analyse using other LLMs.
#ai #cv
In the album some examples how it works. Next we can OCR these images and analyse using other LLMs.
#ai #cv
Also, yesterday's talk at UDS helped to discover KeyBERT's idea of KeyLLM.
Released simplified code to do it using GPT-4o models:
https://github.com/egorsmkv/keyllm
It's useful when you need keywords about some documents (clustering or tagging).
#ai #nlp #llm
Released simplified code to do it using GPT-4o models:
https://github.com/egorsmkv/keyllm
It's useful when you need keywords about some documents (clustering or tagging).
#ai #nlp #llm