π° Anthropic Research - Anthropic Economic Index report: Cadences
https://www.anthropic.com/research/economic-index-june-2026-report
https://www.anthropic.com/research/economic-index-june-2026-report
Anthropic
Anthropic Economic Index report: Cadences
In the latest Anthropic Economic Index report, we look at when people come to Claude, what they produce with it, and how they perceive AIβs impact on their work.
π° NVIDIA - Deploy a Production-Ready NVIDIA AI-Q Blueprint on Oracle Cloud Infrastructure
AI agents have changed a lot in the last two years. The first could only answer one question at a time. Then came multi-turn chat, where the model could keepβ¦
https://developer.nvidia.com/blog/deploy-a-production-ready-nvidia-ai-q-blueprint-on-oracle-cloud-infrastructure/
π° NVIDIA - Creating the NVIDIA Nemotron 3 Ultra NVFP4 Checkpoint with NVIDIA Model Optimizer
As context windows grow longer, moving large model weights efficiently becomes critical to performance. A common way to address this is quantizationβ¦
https://developer.nvidia.com/blog/creating-the-nvidia-nemotron-3-ultra-nvfp4-checkpoint-with-nvidia-model-optimizer/
AI agents have changed a lot in the last two years. The first could only answer one question at a time. Then came multi-turn chat, where the model could keepβ¦
https://developer.nvidia.com/blog/deploy-a-production-ready-nvidia-ai-q-blueprint-on-oracle-cloud-infrastructure/
π° NVIDIA - Creating the NVIDIA Nemotron 3 Ultra NVFP4 Checkpoint with NVIDIA Model Optimizer
As context windows grow longer, moving large model weights efficiently becomes critical to performance. A common way to address this is quantizationβ¦
https://developer.nvidia.com/blog/creating-the-nvidia-nemotron-3-ultra-nvfp4-checkpoint-with-nvidia-model-optimizer/
NVIDIA Technical Blog
Deploy a Production-Ready NVIDIA AI-Q Blueprint on Oracle Cloud Infrastructure
AI agents have changed a lot in the last two years. The first could only answer one question at a time. Then came multi-turn chat, where the model could keep some context across a session. Todayβ¦
π [GitHub Releases] sgl-project/sglang - v0.5.14
https://github.com/sgl-project/sglang/releases/tag/v0.5.14
https://github.com/sgl-project/sglang/releases/tag/v0.5.14
GitHub
Release v0.5.14 Β· sgl-project/sglang
Highlights
New Model Support: GLM-5.2, LiquidAI LFM2.5, Kimi-K2.7-Code, Poolside Laguna-M.1, DiffusionGemma, Zyphra ZAYA1, MiMo-V2-ASR
DeepSeek-V4 on GB300 since Day 0: 5x higher throughput at the ...
New Model Support: GLM-5.2, LiquidAI LFM2.5, Kimi-K2.7-Code, Poolside Laguna-M.1, DiffusionGemma, Zyphra ZAYA1, MiMo-V2-ASR
DeepSeek-V4 on GB300 since Day 0: 5x higher throughput at the ...
π [HF Models] deepseek-ai - DeepSeek-V4-Pro-DSpark
https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro-DSpark
π [HF Models] deepseek-ai - DeepSeek-V4-Flash-DSpark
https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-DSpark
https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro-DSpark
π [HF Models] deepseek-ai - DeepSeek-V4-Flash-DSpark
https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-DSpark
ποΈ Weekly GitHub Activity
π¦ llama.cpp
β Release: b9743 β b9828
β 85 commits
- Added support for Step 3.5/3.7 flash MTP3 speculative decoding #24340,
Granite Speech Plus #24818,
LFM2.5-ColBERT-350M/Embedding-350M #24913,
Eagle3 Qwen3 draft models #24977,
Unlimited-OCR #24969
- Added SSE Replay Buffer to server and UI, allowing text generation to survive HTTP disconnects and resume seamlessly #23226
- Introduced real-time model loading progress tracking via SSE in both server and UI #24828 #24878
- Configured server to create checkpoints before every user message to improve session recovery #24176
- Redesigned the WebUI with a new logo, navigation cleanup, and significant mobile layout improvements #24897
- Enabled dual-GPU tensor parallelism on the SYCL backend via split-mode tensor #24152
- Overhauled Hexagon matrix multiplication kernels with tiled layouts, HVX/HMX microkernels, and graph caching #24954
- Upgraded OpenCL Flash Attention kernels for F16, F32, Q4_0, and Q8_0 #25069
- Moved server model downloading to a dedicated child process #24834
- Added CUDA fast path for strided 2D copies using cudaMemcpy2DAsync #25057
- Reduced synchronization overhead between CPU and CUDA async copies during split compute #20793
- Added 3D convolution support to Vulkan #24612
- Fixed CUDA integer overflows and transposed copy failures #24706 #25000
- Fixed incorrect vector dot computations on SVE-enabled ARM CPUs #24699
π All changes | Latest release
π¨ stable-diffusion.cpp
β Release: master-709-92a3b73 β master-721-8caa3f9
β 12 commits
- Added support for Boogu image generation #1688
- Added support for Krea2 models #1705
- Introduced guidance_schedule support for generation control #1684
- Added logit-normal scheduler #1669
- Added --eager-load flag to pre-load parameters during model initialization #1687
- Added --prompt-file and --negative-prompt-file flags for file-based inputs #1693
- Fixed memory mapping by avoiding writable mmap for read-only weights #1698
π All changes | Latest release
π€ Fresh models trending on HuggingFace:
empero-ai/Qwythos-9B-Claude-Mythos-5-1M β‘488
krea/Krea-2-Turbo β‘310
krea/Krea-2-Raw β‘214
deepreinforce-ai/Ornith-1.0-9B β‘167
deepreinforce-ai/Ornith-1.0-35B β‘161
deepreinforce-ai/Ornith-1.0-397B β‘121
Chunjiang-Intelligence/DeepSeek-v4-Fable β‘112
hustvl/Moebius β‘51
AutoArk-AI/ARK-ASR-3B β‘37
paom/texture2albedo-v2 β‘32
SupraLabs/Supra-A2A-Nano-Exp β‘30
Gryphe/Gemma-4-26B-A4B-StyleTune-V2 β‘24
0xSero/GLM-5.2-504B β‘19
g-astruc/UniverSat β‘18
allenai/tmax-27b β‘18
ValiantLabs/Qwen3.6-27B-Esper4 β‘14
wikeeyang/Flux2-Klein-9B-True-V3 β‘14
vrgamedevgirl84/Krea2_Enhancer β‘12
π¦ llama.cpp
β Release: b9743 β b9828
β 85 commits
- Added support for Step 3.5/3.7 flash MTP3 speculative decoding #24340,
Granite Speech Plus #24818,
LFM2.5-ColBERT-350M/Embedding-350M #24913,
Eagle3 Qwen3 draft models #24977,
Unlimited-OCR #24969
- Added SSE Replay Buffer to server and UI, allowing text generation to survive HTTP disconnects and resume seamlessly #23226
- Introduced real-time model loading progress tracking via SSE in both server and UI #24828 #24878
- Configured server to create checkpoints before every user message to improve session recovery #24176
- Redesigned the WebUI with a new logo, navigation cleanup, and significant mobile layout improvements #24897
- Enabled dual-GPU tensor parallelism on the SYCL backend via split-mode tensor #24152
- Overhauled Hexagon matrix multiplication kernels with tiled layouts, HVX/HMX microkernels, and graph caching #24954
- Upgraded OpenCL Flash Attention kernels for F16, F32, Q4_0, and Q8_0 #25069
- Moved server model downloading to a dedicated child process #24834
- Added CUDA fast path for strided 2D copies using cudaMemcpy2DAsync #25057
- Reduced synchronization overhead between CPU and CUDA async copies during split compute #20793
- Added 3D convolution support to Vulkan #24612
- Fixed CUDA integer overflows and transposed copy failures #24706 #25000
- Fixed incorrect vector dot computations on SVE-enabled ARM CPUs #24699
π All changes | Latest release
π¨ stable-diffusion.cpp
β Release: master-709-92a3b73 β master-721-8caa3f9
β 12 commits
- Added support for Boogu image generation #1688
- Added support for Krea2 models #1705
- Introduced guidance_schedule support for generation control #1684
- Added logit-normal scheduler #1669
- Added --eager-load flag to pre-load parameters during model initialization #1687
- Added --prompt-file and --negative-prompt-file flags for file-based inputs #1693
- Fixed memory mapping by avoiding writable mmap for read-only weights #1698
π All changes | Latest release
π€ Fresh models trending on HuggingFace:
empero-ai/Qwythos-9B-Claude-Mythos-5-1M β‘488
krea/Krea-2-Turbo β‘310
krea/Krea-2-Raw β‘214
deepreinforce-ai/Ornith-1.0-9B β‘167
deepreinforce-ai/Ornith-1.0-35B β‘161
deepreinforce-ai/Ornith-1.0-397B β‘121
Chunjiang-Intelligence/DeepSeek-v4-Fable β‘112
hustvl/Moebius β‘51
AutoArk-AI/ARK-ASR-3B β‘37
paom/texture2albedo-v2 β‘32
SupraLabs/Supra-A2A-Nano-Exp β‘30
Gryphe/Gemma-4-26B-A4B-StyleTune-V2 β‘24
0xSero/GLM-5.2-504B β‘19
g-astruc/UniverSat β‘18
allenai/tmax-27b β‘18
ValiantLabs/Qwen3.6-27B-Esper4 β‘14
wikeeyang/Flux2-Klein-9B-True-V3 β‘14
vrgamedevgirl84/Krea2_Enhancer β‘12
GitHub
Support Step3.5/3.7 flash mtp3 by forforever73 Β· Pull Request #24340 Β· ggml-org/llama.cpp
Overview
follow-up to #23274.(cc @pwilkin )
π Full data-flow trace β couldn't think of a good way to draw this, so I wrote it all down instead. It's long, but every byte is load-be...
follow-up to #23274.(cc @pwilkin )
π Full data-flow trace β couldn't think of a good way to draw this, so I wrote it all down instead. It's long, but every byte is load-be...
β€2
π [HF Models] deepseek-ai - eagle3_gemma4_12b_ttt7
https://huggingface.co/deepseek-ai/eagle3_gemma4_12b_ttt7
π [HF Models] deepseek-ai - eagle3_qwen3_14b_ttt7
https://huggingface.co/deepseek-ai/eagle3_qwen3_14b_ttt7
π [HF Models] deepseek-ai - eagle3_qwen3_8b_ttt7
https://huggingface.co/deepseek-ai/eagle3_qwen3_8b_ttt7
π [HF Models] deepseek-ai - eagle3_qwen3_4b_ttt7
https://huggingface.co/deepseek-ai/eagle3_qwen3_4b_ttt7
π [HF Models] deepseek-ai - dflash_gemma4_12b_block7
https://huggingface.co/deepseek-ai/dflash_gemma4_12b_block7
π [HF Models] deepseek-ai - dflash_qwen3_14b_block7
https://huggingface.co/deepseek-ai/dflash_qwen3_14b_block7
π [HF Models] deepseek-ai - dflash_qwen3_8b_block7
https://huggingface.co/deepseek-ai/dflash_qwen3_8b_block7
π [HF Models] deepseek-ai - dflash_qwen3_4b_block7
https://huggingface.co/deepseek-ai/dflash_qwen3_4b_block7
π [HF Models] deepseek-ai - dspark_gemma4_12b_block7
https://huggingface.co/deepseek-ai/dspark_gemma4_12b_block7
π [HF Models] deepseek-ai - dspark_qwen3_14b_block7
https://huggingface.co/deepseek-ai/dspark_qwen3_14b_block7
π [HF Models] deepseek-ai - dspark_qwen3_8b_block7
https://huggingface.co/deepseek-ai/dspark_qwen3_8b_block7
π [HF Models] deepseek-ai - dspark_qwen3_4b_block7
https://huggingface.co/deepseek-ai/dspark_qwen3_4b_block7
https://huggingface.co/deepseek-ai/eagle3_gemma4_12b_ttt7
π [HF Models] deepseek-ai - eagle3_qwen3_14b_ttt7
https://huggingface.co/deepseek-ai/eagle3_qwen3_14b_ttt7
π [HF Models] deepseek-ai - eagle3_qwen3_8b_ttt7
https://huggingface.co/deepseek-ai/eagle3_qwen3_8b_ttt7
π [HF Models] deepseek-ai - eagle3_qwen3_4b_ttt7
https://huggingface.co/deepseek-ai/eagle3_qwen3_4b_ttt7
π [HF Models] deepseek-ai - dflash_gemma4_12b_block7
https://huggingface.co/deepseek-ai/dflash_gemma4_12b_block7
π [HF Models] deepseek-ai - dflash_qwen3_14b_block7
https://huggingface.co/deepseek-ai/dflash_qwen3_14b_block7
π [HF Models] deepseek-ai - dflash_qwen3_8b_block7
https://huggingface.co/deepseek-ai/dflash_qwen3_8b_block7
π [HF Models] deepseek-ai - dflash_qwen3_4b_block7
https://huggingface.co/deepseek-ai/dflash_qwen3_4b_block7
π [HF Models] deepseek-ai - dspark_gemma4_12b_block7
https://huggingface.co/deepseek-ai/dspark_gemma4_12b_block7
π [HF Models] deepseek-ai - dspark_qwen3_14b_block7
https://huggingface.co/deepseek-ai/dspark_qwen3_14b_block7
π [HF Models] deepseek-ai - dspark_qwen3_8b_block7
https://huggingface.co/deepseek-ai/dspark_qwen3_8b_block7
π [HF Models] deepseek-ai - dspark_qwen3_4b_block7
https://huggingface.co/deepseek-ai/dspark_qwen3_4b_block7
huggingface.co
deepseek-ai/eagle3_gemma4_12b_ttt7 Β· Hugging Face
Weβre on a journey to advance and democratize artificial intelligence through open source and open science.
π° PyTorch - Introducing Cross-Repository CI Relay: Scalable CI for PyTorchβs Out-of-Tree Backends
TL;DR PyTorch now has a Cross-Repository CI Relay (CRCR) that automatically triggers and tracks CI in downstream repositories whenever a PR is opened or a commit is pushed against pytorch/pytorch....
https://pytorch.org/blog/introducing-cross-repository-ci-relay-scalable-ci-for-pytorchs-out-of-tree-backends/
TL;DR PyTorch now has a Cross-Repository CI Relay (CRCR) that automatically triggers and tracks CI in downstream repositories whenever a PR is opened or a commit is pushed against pytorch/pytorch....
https://pytorch.org/blog/introducing-cross-repository-ci-relay-scalable-ci-for-pytorchs-out-of-tree-backends/
β€1
π° HuggingFace - DiScoFormer: One transformer for density and score, across distributions
https://huggingface.co/blog/allenai/discoformer
https://huggingface.co/blog/allenai/discoformer
huggingface.co
DiScoFormer: One transformer for density and score, across distributions
A Blog post by Ai2 on Hugging Face
π [GitHub Releases] vllm-project/vllm - v0.24.0
https://github.com/vllm-project/vllm/releases/tag/v0.24.0
https://github.com/vllm-project/vllm/releases/tag/v0.24.0
GitHub
Release v0.24.0 Β· vllm-project/vllm
vLLM v0.24.0 Release Notes
Highlights
This release features 571 commits from 256 contributors (77 new)!
MiniMax-M3: Added support for the new MiniMax-M3 model (#45381), with a fast follow-on of BF...
Highlights
This release features 571 commits from 256 contributors (77 new)!
MiniMax-M3: Added support for the new MiniMax-M3 model (#45381), with a fast follow-on of BF...
β€1
π [GitHub Releases] open-webui/open-webui - v0.10.1
https://github.com/open-webui/open-webui/releases/tag/v0.10.1
https://github.com/open-webui/open-webui/releases/tag/v0.10.1
GitHub
Release v0.10.1 Β· open-webui/open-webui
Fixed
π€ Shared folder read-only chats no longer sign users out. Opening or reading chats from shared folders now keeps the current session active when a resource-level access error is returned, in...
π€ Shared folder read-only chats no longer sign users out. Opening or reading chats from shared folders now keeps the current session active when a resource-level access error is returned, in...
π° OpenAI - Mapping Europeβs AI Workforce Opportunity
https://openai.com/index/mapping-ai-jobs-transition-eu
π° OpenAI - HP Inc. launches Frontier strategic partnership with OpenAI
https://openai.com/index/hp-frontier-partnership
https://openai.com/index/mapping-ai-jobs-transition-eu
π° OpenAI - HP Inc. launches Frontier strategic partnership with OpenAI
https://openai.com/index/hp-frontier-partnership
OpenAI
Mapping Europeβs AI Workforce Opportunity
A new OpenAI report maps how AI could reshape jobs across the EU, highlighting which occupations may face automation, growth, or workflow changes.
π° NVIDIA - How to Govern Autonomous Agents in Enterprise AI Factories
AI agents are quickly moving beyond chat. They inspect code, run tests, read documents, search knowledge bases, query internal systems, and operate for hours onβ¦
https://developer.nvidia.com/blog/how-to-govern-autonomous-agents-in-enterprise-ai-factories/
AI agents are quickly moving beyond chat. They inspect code, run tests, read documents, search knowledge bases, query internal systems, and operate for hours onβ¦
https://developer.nvidia.com/blog/how-to-govern-autonomous-agents-in-enterprise-ai-factories/
NVIDIA Technical Blog
How to Govern Autonomous Agents in Enterprise AI Factories
AI agents are quickly moving beyond chat. They inspect code, run tests, read documents, search knowledge bases, query internal systems, and operate for hours on behalf of a user.
π [HF Models] microsoft - vermeer-XL-CA
https://huggingface.co/microsoft/vermeer-XL-CA
π [HF Models] microsoft - Dayhoff-170M-GRS-SS-74000
https://huggingface.co/microsoft/Dayhoff-170M-GRS-SS-74000
π [HF Models] microsoft - Dayhoff-170M-GRS-SS-86000
https://huggingface.co/microsoft/Dayhoff-170M-GRS-SS-86000
https://huggingface.co/microsoft/vermeer-XL-CA
π [HF Models] microsoft - Dayhoff-170M-GRS-SS-74000
https://huggingface.co/microsoft/Dayhoff-170M-GRS-SS-74000
π [HF Models] microsoft - Dayhoff-170M-GRS-SS-86000
https://huggingface.co/microsoft/Dayhoff-170M-GRS-SS-86000
huggingface.co
microsoft/vermeer-XL-CA Β· Hugging Face
Weβre on a journey to advance and democratize artificial intelligence through open source and open science.
π [GitHub Releases] invoke-ai/InvokeAI - InvokeAI v6.13.5 (release candidate 1)
https://github.com/invoke-ai/InvokeAI/releases/tag/v6.13.5.rc1
https://github.com/invoke-ai/InvokeAI/releases/tag/v6.13.5.rc1
GitHub
Release InvokeAI v6.13.5 (release candidate 1) Β· invoke-ai/InvokeAI
This is a maintenance release of InvokeAI focused on bug fixes and stability. Version 6.14.0 will be the next major feature release, featuring video generation, multiple GPU support, the Wan 2.2 im...
π° HuggingFace - Featuring Every Eval Ever Results on Hugging Face Model Pages
https://huggingface.co/blog/eee-community-evals
https://huggingface.co/blog/eee-community-evals
huggingface.co
Featuring Every Eval Ever Results on Hugging Face Model Pages
Weβre on a journey to advance and democratize artificial intelligence through open source and open science.
π [HF Models] microsoft - GELab-Zero-4B-preview-Sico-Evolution
https://huggingface.co/microsoft/GELab-Zero-4B-preview-Sico-Evolution
https://huggingface.co/microsoft/GELab-Zero-4B-preview-Sico-Evolution
huggingface.co
microsoft/GELab-Zero-4B-preview-Sico-Evolution Β· Hugging Face
Weβre on a journey to advance and democratize artificial intelligence through open source and open science.
π° HuggingFace - Why Specialization Is Inevitable
https://huggingface.co/blog/Dharma-AI/why-specialization-is-inevitable
https://huggingface.co/blog/Dharma-AI/why-specialization-is-inevitable
huggingface.co
Why Specialization Is Inevitable
A Blog post by Dharma-AI on Hugging Face
π [HF Models] google - tabfm-1.0.0-jax
https://huggingface.co/google/tabfm-1.0.0-jax
π [HF Models] google - tabfm-1.0.0-pytorch
https://huggingface.co/google/tabfm-1.0.0-pytorch
https://huggingface.co/google/tabfm-1.0.0-jax
π [HF Models] google - tabfm-1.0.0-pytorch
https://huggingface.co/google/tabfm-1.0.0-pytorch
huggingface.co
google/tabfm-1.0.0-jax Β· Hugging Face
Weβre on a journey to advance and democratize artificial intelligence through open source and open science.