OpenSourcePulse
2.69K subscribers
1.12K photos
1.12K links
Download Telegram
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10213
πŸ“… Released: 2026-07-31 20:10 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

Support rotated kv cache quant (#26180)

</details>

Website:
- <https://llama.app>

macOS/iOS:
- macOS Apple Silicon (arm64)
- macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED
- macOS Intel (x64)
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10214
πŸ“… Released: 2026-07-31 20:46 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

mtmd: add nembdhead (#26342)

Co-authored-by: Daniel Han <unslothai@gmail.com>

</details>

Website:
- <https://llama.app>

macOS/iOS:
- macOS Apple Silicon (arm64)
- macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10216
πŸ“… Released: 2026-07-31 22:04 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

vulkan: add POOL1D op (#25431)

* vulkan : add pool1d push constants and pipeline field

Declared data structures needed for POOL1D OP, which are the vk
oppool1dpushconstants struct and pipelinepool1df32 field.

* vulkan : add pool1d compute shader

Added pool1d.comp for Vulkan backend mirroring the existing pool2d shader.

* vulkan : add full GGML
OPPOOL1D support

Adde...

πŸ”— View on GitHub
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10217
πŸ“… Released: 2026-08-01 06:50 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

chat : enable tool call in thinking for DS4 (#26269)

</details>

Website:
- <https://llama.app>

macOS/iOS:
- macOS Apple Silicon (arm64)
- macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED
- macOS Intel (x64)
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10218
πŸ“… Released: 2026-08-01 12:46 UTC
πŸ“Š Assets: 24 files

πŸ“ What's New:
<details open>

mtmd: add minicpmv46 downsample (#25993)

add minicpmv46 downsample

Signed-off-by: tc-mb <
tianchi_cai@icloud.com>

put downsample mode inside gguf.

Signed-off-by: tc-mb <tianchicai@icloud.com>

* build mtmd
imagepreprocessorllavauhd

Signed-off-by: tc-mb <
tianchicai@icloud.com>

fix code

Signed-off-by: tc-mb <
tianchi_cai@icloud.com>

add convert

Signed-off-by: t...

πŸ”— View on GitHub
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10219
πŸ“… Released: 2026-08-01 16:46 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

cli : persist reasoningcontent in chat history (#26362)

* cli : persist reasoning
content in chat history

llama-cli collected reasoning from the stream for display but only
stored assistant content in messages, so --reasoning-preserve could
not re-inject prior thoughts on later turns.

</details>

Website:
- <https://llama.app>

macOS/iOS:
- macOS Apple Silicon (arm6...

πŸ”— [View on GitHub

πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10221
πŸ“… Released: 2026-08-01 19:30 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

vendor : update BoringSSL to 0.20260730.0 (#26353)

</details>

Website:
- <https://llama.app>

macOS/iOS:
- macOS Apple Silicon (arm64)
- macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED
- macOS Intel (x64)
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10223
πŸ“… Released: 2026-08-01 22:48 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

test: fix some CI errors (#26415)

</details>

Website:
- <https://llama.app>

macOS/iOS:
- macOS Apple Silicon (arm64)
- macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED
- macOS Intel (x64)
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10224
πŸ“… Released: 2026-08-02 07:03 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

ggml-webgpu: add support for f16 repeat (#26307)

</details>

Website:
- <https://llama.app>

macOS/iOS:
- macOS Apple Silicon (arm64)
- macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED
- macOS Intel (x64)
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10226
πŸ“… Released: 2026-08-02 08:15 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

sycl: fix classification of iGPUs (#26105)

</details>

Website:
- <https://llama.app>

macOS/iOS:
- macOS Apple Silicon (arm64)
- macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED
- macOS Intel (x64)
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10225
πŸ“… Released: 2026-08-02 08:30 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

model : load MiMo V2 MTP tensors only if used (#26412)

</details>

Website:
- <https://llama.app>

macOS/iOS:
- macOS Apple Silicon (arm64)
- macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED
- macOS Intel (x64)...

πŸ”— View on GitHub
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10227
πŸ“… Released: 2026-08-02 09:43 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

chat : add qwen3 specialized parser (#26252)

Add tagged thinking tool parser

chat : refactor and add permute helper

cont : add support for <tool_call> omission

cont : update tool delimiters

cont : add comment for qwen3-coder

cont : fix trigger pattern for <function

---------

Co-authored-by: Bart de Boer <bart.deboer@gmail.com>

</details>

Website:
- <htt...

πŸ”— View on GitHub
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ thefeed
🏷 Category: Community Suggested
πŸ”– Version: v0.38.0
πŸ“… Released: 2026-08-02 11:01 UTC
πŸ“Š Assets: 27 files

πŸ“ What's New:
## Install

![Google Play](https://play.google.com/store/apps/details?id=com.thefeed.android)
![TestFlight](https://testflight.apple.com/join/J6bfxDdZ)

Android and iOS install...

πŸ”— View on GitHub
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10228
πŸ“… Released: 2026-08-02 13:28 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

DeepseekV4 MTP + DSpark (#25784)

</details>

Website:
- <https://llama.app>

macOS/iOS:
- macOS Apple Silicon (arm64)
- macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED
- macOS Intel (x64)
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10229
πŸ“… Released: 2026-08-02 14:31 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

opencl: bugfix increment refcount in ggmlbackendopenclinit() (#26162)

Incrementing ref_count at the beginning is important later
in the free() method of the ggml_backend_opencl_context at program end.
If we do not increment the ref_count, the result would be -1 here,
and consequently, the profiling data would not be flushed and written.
( #ifdef GGMLOPENCLPROFILI...

πŸ”— View on GitHub
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10231
πŸ“… Released: 2026-08-02 18:14 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

common: support the DSpark sidecar resolution (#26458)

The dspark- files resolve like the other speculative sidecars: the
-hfd tag applies to them, a requested sidecar resolves without a full
model at the tag, and an explicit -md selection disables the discovery.
When no type is requested, dspark outranks dflash in the auto-selection
since its sidecar carries the extra Markov h...

πŸ”— View on GitHub
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10232
πŸ“… Released: 2026-08-02 18:57 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

metal: implement DeepSeek V4 hyper-connections (#26459)

- Implement GGMLOPDSV4HCCOMB, GGMLOPDSV4HCPRE, and
GGMLOPDSV4HCPOST with SIMDgroup register and shuffle optimized kernels.
- Add Metal dispatch and support plumbing and test the production Sinkhorn
iteration count and embedding width.

Assisted-by: Codex

Co-authored-by: Thiago Padilha <thiago@padilha.cc>

...

πŸ”— View on GitHub
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10234
πŸ“… Released: 2026-08-02 20:17 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

metal : add F16 support for bin ops (#26465)

</details>

Website:
- <https://llama.app>

macOS/iOS:
- macOS Apple Silicon (arm64)
- macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED
- macOS Intel (x64)
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10235
πŸ“… Released: 2026-08-02 21:02 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

metal : add SILUBACK (#25982)

* feat(silu
back): implemented siluback op for f32

* fix(silu
back): removed redundant asserts in ggml-metal-ops.cpp function ggmlmetalopsiluback.

</details>

Website:
- <https://llama.app>

macOS/iOS:
- macOS Apple Silicon (arm64)
...

πŸ”— View on GitHub
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10237
πŸ“… Released: 2026-08-03 07:50 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

llama : MTP support for DeepSeek V3.2 (#26457)

llama : MTP support for DeepSeek V3.2

model : no need to include MTP layers during DeepSeek V3.2 model type discovery

---------

Co-authored-by: StanisΕ‚aw Szymczyk <sszymczy@gmail.com>

</details>

Website:
- <https://llama.app>

macOS/iOS:
- macOS Apple Silicon (arm64)
πŸ“₯ Download Release
πŸš€ New Release Alert!

πŸ“¦ llama.cpp
🏷 Category: AI
πŸ”– Version: b10238
πŸ“… Released: 2026-08-03 08:56 UTC
πŸ“Š Assets: 25 files

πŸ“ What's New:
<details open>

model: MTP support for Qwen3-Next (#25589)

mtp for qwen3nex

fix for python type-check

Fix to compute num_mtp from directly mtp layer

define optnummtplayers in QwenMtpMixin and fix some comments

Fix for python type check

Update gguf-py/gguf/constants.py

Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@huggingface.co>

rebase and add load_mtp flags

U...

πŸ”— View on GitHub
πŸ“₯ Download Release