Fine-tuned Whisper model for Multilingual Speech-to-Phonemes recognition:
https://colab.research.google.com/drive/1cJD_yQMM71IaLxkC0mEgYLOb2DX7p_LA?usp=sharing
#audio #ai
https://colab.research.google.com/drive/1cJD_yQMM71IaLxkC0mEgYLOb2DX7p_LA?usp=sharing
#audio #ai
Google
Speech-to-Phonemes, test with Whisper-PPT.ipynb
Colab notebook
Made a simple Containerfile to run llamafile using podman/docker:
https://github.com/crs-org/multilingual-pii-detection/blob/main/containers/llamafile/Containerfile
Size of the image: 6.68 GB
#ai #llm
https://github.com/crs-org/multilingual-pii-detection/blob/main/containers/llamafile/Containerfile
Size of the image: 6.68 GB
#ai #llm
GitHub
multilingual-pii-detection/containers/llamafile/Containerfile at main Β· crs-org/multilingual-pii-detection
An API for the Personally Identifiable Information detection task - crs-org/multilingual-pii-detection
I think I've seen it before but let's write here as well about this model - https://huggingface.co/osmosis-ai/Osmosis-Structure-0.6B
This model gives an ability to extract information in text using JSON schema
Quantized model (8-bit) requires about 1602MiB of GPU
#ai #llm
This model gives an ability to extract information in text using JSON schema
Quantized model (8-bit) requires about 1602MiB of GPU
#ai #llm
π2π€―1
Forwarded from Hacker News