Alex
https://www.cerebras.ai/blog/introducing-cerebras-code
Other coding models:
- CodeGemma
- Codestral
- Code LLaMA (there variants: Python, instruct, base)
- StarCoder
- DeepSeek Coder
- Mellum (Python, Kotline, Base)
- GLM 4.5
- CodeGeeX
- CodeGemma
- Codestral
- Code LLaMA (there variants: Python, instruct, base)
- StarCoder
- DeepSeek Coder
- Mellum (Python, Kotline, Base)
- GLM 4.5
- CodeGeeX
https://github.com/EricLBuehler/mistral.rs
Wanna try it to do local inference with Gemma models
Build with:
#rust
Wanna try it to do local inference with Gemma models
Build with:
MISTRALRS_METAL_PRECOMPILE=0 cargo build --release --features "metal accelerate"
#rust
doing something
https://github.com/EricLBuehler/mistral.rs Wanna try it to do local inference with Gemma models Build with: MISTRALRS_METAL_PRECOMPILE=0 cargo build --release --features "metal accelerate" #rust
It has Python bindings with useful things
https://github.com/EricLBuehler/mistral.rs/blob/master/examples/python/mcp_client.py
https://github.com/EricLBuehler/mistral.rs/blob/master/examples/python/mcp_client.py
Hugging Face has a page with available providers and model list they serve
https://huggingface.co/inference/models
It shows the speed in tokens/s so you can choose one for faster inference
https://huggingface.co/inference/models
It shows the speed in tokens/s so you can choose one for faster inference
doing something
Hugging Face has a page with available providers and model list they serve https://huggingface.co/inference/models It shows the speed in tokens/s so you can choose one for faster inference
I've chose LLaMA to check grammar mistakes in my codebase
Time is 1.5 seconds for everything
Time is 1.5 seconds for everything
❤1