📰 Google AI Blog - How A2A is Building a World of Collaborative Agents
Celebrating the first anniversary of the Agent-to-Agent (A2A) protocol, this blog post highlights how the framework enables autonomous AI agents to securely collaborate and hand off tasks without the rigidity of traditional APIs. By delegating complex workflows to specialized peer agents, A2A prevents context pollution, ensures data privacy, and simplifies application design through modularity. To demonstrate this ecosystem in action, the post spotlights FoldRun—an agentic interface for life sciences that orchestrates complex protein structure predictions—alongside diverse A2A use cases spanning commerce, data streaming, DevOps, and telecommunications.
https://developers.googleblog.com/en/how-a2a-is-building-a-world-of-collaborative-agents/
Celebrating the first anniversary of the Agent-to-Agent (A2A) protocol, this blog post highlights how the framework enables autonomous AI agents to securely collaborate and hand off tasks without the rigidity of traditional APIs. By delegating complex workflows to specialized peer agents, A2A prevents context pollution, ensures data privacy, and simplifies application design through modularity. To demonstrate this ecosystem in action, the post spotlights FoldRun—an agentic interface for life sciences that orchestrates complex protein structure predictions—alongside diverse A2A use cases spanning commerce, data streaming, DevOps, and telecommunications.
https://developers.googleblog.com/en/how-a2a-is-building-a-world-of-collaborative-agents/
Googleblog
Google for Developers Blog - News about Web, Mobile, AI and Cloud
Discover how the Agent-to-Agent (A2A) protocol is shifting AI from isolated tools to a collaborative ecosystem, enabling secure, autonomous agent handoffs and scalable workflows like FoldRun.
📰 Claude Blog - Steering Claude Code: CLAUDE.md files, skills, hooks, rules, subagents and more
https://claude.com/blog/steering-claude-code-skills-hooks-rules-subagents-and-more
📰 Claude Blog - Centrally manage authorization for MCP connectors
https://claude.com/blog/enterprise-managed-auth
📰 Claude Blog - Claude Code now supports artifacts
https://claude.com/blog/artifacts-in-claude-code
📰 Claude Blog - Meet the winners of our Claude Opus 4.8 Build Day hackathon
https://claude.com/blog/meet-the-winners-of-our-claude-opus-4-8-build-day-hackathon
📰 Claude Blog - Claude Design now stays on brand for daily work
https://claude.com/blog/claude-design-stays-on-brand-for-daily-work
📰 Claude Blog - Secure access to the Claude Platform with Workload Identity Federation
https://claude.com/blog/workload-identity-federation
https://claude.com/blog/steering-claude-code-skills-hooks-rules-subagents-and-more
📰 Claude Blog - Centrally manage authorization for MCP connectors
https://claude.com/blog/enterprise-managed-auth
📰 Claude Blog - Claude Code now supports artifacts
https://claude.com/blog/artifacts-in-claude-code
📰 Claude Blog - Meet the winners of our Claude Opus 4.8 Build Day hackathon
https://claude.com/blog/meet-the-winners-of-our-claude-opus-4-8-build-day-hackathon
📰 Claude Blog - Claude Design now stays on brand for daily work
https://claude.com/blog/claude-design-stays-on-brand-for-daily-work
📰 Claude Blog - Secure access to the Claude Platform with Workload Identity Federation
https://claude.com/blog/workload-identity-federation
Claude
Steering Claude Code: when to use CLAUDE.md, skills, hooks, and subagents | Claude by Anthropic
Seven ways to steer Claude Code—CLAUDE.md files, rules, skills, subagents, hooks, and more—and when to use each, based on context cost and authority.
🆕 [HF Models] FunAudioLLM - Fun-ASR-Nano-GGUF
https://huggingface.co/FunAudioLLM/Fun-ASR-Nano-GGUF
🆕 [HF Models] FunAudioLLM - Paraformer-GGUF
https://huggingface.co/FunAudioLLM/Paraformer-GGUF
🆕 [HF Models] FunAudioLLM - SenseVoiceSmall-GGUF
https://huggingface.co/FunAudioLLM/SenseVoiceSmall-GGUF
🆕 [HF Models] FunAudioLLM - fsmn-vad-GGUF
https://huggingface.co/FunAudioLLM/fsmn-vad-GGUF
https://huggingface.co/FunAudioLLM/Fun-ASR-Nano-GGUF
🆕 [HF Models] FunAudioLLM - Paraformer-GGUF
https://huggingface.co/FunAudioLLM/Paraformer-GGUF
🆕 [HF Models] FunAudioLLM - SenseVoiceSmall-GGUF
https://huggingface.co/FunAudioLLM/SenseVoiceSmall-GGUF
🆕 [HF Models] FunAudioLLM - fsmn-vad-GGUF
https://huggingface.co/FunAudioLLM/fsmn-vad-GGUF
🗓️ Weekly GitHub Activity
🦙 llama.cpp
└ Release: b9627 → b9743
└ 116 commits
- Added support for Cohere2MoE (North Code / Tiny Aya) and GLM-5.2 models. #24615 #24770
- Integrated Eagle3 speculative decoding support for Qwen 3.5 and 3.6. #24593
- Updated OpenVINO backend to 2026.2 with context-shift, Q5_1 weights, and Gemma 4 support. #24503
- Introduced a model management API to the server router for remote model downloads and deletion. #23976
- UI improvements: added HEIC/HEIF image support, SVG/Mermaid rendering with source toggles, and markdown rendering for thinking blocks. #24137 #24080 #24611
- Optimized AMX performance on CPU and improved i-quants prefill speeds for WebGPU. #24806 #24530
- Enhanced Metal backend with concat support for F16/BF16 and rope_back operator. #24724 #24725
- Fixed significant whitespace issues in chat grammar generation and double-escaping in tool-call parsing. #24624 #24667
- Server now includes real-time generation speed metrics and JSONL conversation exports. #24291 #24688
- SYCL backend updates: added Conv2D/Conv3D support, dev-to-dev memcpy, and set F16 as default. #24600 #24476 #23996
🔗 All changes | Latest release
🎨 stable-diffusion.cpp
└ Release: master-694-276025e → master-709-92a3b73
└ 15 commits
- Added RPC support for remote compute execution #1629
- Implemented PuLID-Flux identity-injection support for Flux models #1595
- Added support for cancelling ongoing generations with partial image batch returns #1124
- Introduced disk parameters backend support #1651
- Added backend-specific max-VRAM budgets bb90bfa
- Fixed handling of oversized Vulkan parameter tensors #1662
- Synchronized core library with latest GGML #1656
🔗 All changes | Latest release
🤗 Fresh models trending on HuggingFace:
WeiboAI/VibeThinker-3B ♡511
prefeitura-rio/Rio-3.5-Open-397B ♡327
owensong/Inflect-Nano-v1 | gguf ♡140
Zyphra/ZONOS2 ♡118
datalab-to/lift ♡86
poolside/Laguna-M.1 ♡74
Boogu/Boogu-Image-0.1-Edit ♡67
AlexWortega/SIQ-1-35B ♡59
Boogu/Boogu-Image-0.1-Turbo ♡37
SupraLabs/Supra-1.5-50M-Instruct-exp ♡37
Boogu/Boogu-Image-0.1-Base ♡32
HKUSTAudio/AudioX-Turbo ♡29
FINAL-Bench/Darwin-398B-JGOS ♡28
Multilingual-Multimodal-NLP/LoopCoder-V2 ♡26
YTan2000/Qwen3.6-27B-MTP-TQ3_4S ♡16
Danrisi/UltraReal_FineTune_Anima_base1_v3 ♡14
catnip-ai-tech/MaineCoon ♡14
🦙 llama.cpp
└ Release: b9627 → b9743
└ 116 commits
- Added support for Cohere2MoE (North Code / Tiny Aya) and GLM-5.2 models. #24615 #24770
- Integrated Eagle3 speculative decoding support for Qwen 3.5 and 3.6. #24593
- Updated OpenVINO backend to 2026.2 with context-shift, Q5_1 weights, and Gemma 4 support. #24503
- Introduced a model management API to the server router for remote model downloads and deletion. #23976
- UI improvements: added HEIC/HEIF image support, SVG/Mermaid rendering with source toggles, and markdown rendering for thinking blocks. #24137 #24080 #24611
- Optimized AMX performance on CPU and improved i-quants prefill speeds for WebGPU. #24806 #24530
- Enhanced Metal backend with concat support for F16/BF16 and rope_back operator. #24724 #24725
- Fixed significant whitespace issues in chat grammar generation and double-escaping in tool-call parsing. #24624 #24667
- Server now includes real-time generation speed metrics and JSONL conversation exports. #24291 #24688
- SYCL backend updates: added Conv2D/Conv3D support, dev-to-dev memcpy, and set F16 as default. #24600 #24476 #23996
🔗 All changes | Latest release
🎨 stable-diffusion.cpp
└ Release: master-694-276025e → master-709-92a3b73
└ 15 commits
- Added RPC support for remote compute execution #1629
- Implemented PuLID-Flux identity-injection support for Flux models #1595
- Added support for cancelling ongoing generations with partial image batch returns #1124
- Introduced disk parameters backend support #1651
- Added backend-specific max-VRAM budgets bb90bfa
- Fixed handling of oversized Vulkan parameter tensors #1662
- Synchronized core library with latest GGML #1656
🔗 All changes | Latest release
🤗 Fresh models trending on HuggingFace:
WeiboAI/VibeThinker-3B ♡511
prefeitura-rio/Rio-3.5-Open-397B ♡327
owensong/Inflect-Nano-v1 | gguf ♡140
Zyphra/ZONOS2 ♡118
datalab-to/lift ♡86
poolside/Laguna-M.1 ♡74
Boogu/Boogu-Image-0.1-Edit ♡67
AlexWortega/SIQ-1-35B ♡59
Boogu/Boogu-Image-0.1-Turbo ♡37
SupraLabs/Supra-1.5-50M-Instruct-exp ♡37
Boogu/Boogu-Image-0.1-Base ♡32
HKUSTAudio/AudioX-Turbo ♡29
FINAL-Bench/Darwin-398B-JGOS ♡28
Multilingual-Multimodal-NLP/LoopCoder-V2 ♡26
YTan2000/Qwen3.6-27B-MTP-TQ3_4S ♡16
Danrisi/UltraReal_FineTune_Anima_base1_v3 ♡14
catnip-ai-tech/MaineCoon ♡14
GitHub
chat: add dedicated Cohere2MoE (North Code) parser by pwilkin · Pull Request #24615 · ggml-org/llama.cpp
Overview
The Cohere2 MoE template is pretty special, so using the autoparser even with workarounds didn't really work. Needed a dedicated parser.
Additional information
Please use the templ...
The Cohere2 MoE template is pretty special, so using the autoparser even with workarounds didn't really work. Needed a dedicated parser.
Additional information
Please use the templ...
🔓 [HF Models] inclusionAI - Sing-Guard-2b
https://huggingface.co/inclusionAI/Sing-Guard-2b
🔓 [HF Models] inclusionAI - Sing-Guard-8b
https://huggingface.co/inclusionAI/Sing-Guard-8b
🔓 [HF Models] inclusionAI - Sing-Guard-4b
https://huggingface.co/inclusionAI/Sing-Guard-4b
https://huggingface.co/inclusionAI/Sing-Guard-2b
🔓 [HF Models] inclusionAI - Sing-Guard-8b
https://huggingface.co/inclusionAI/Sing-Guard-8b
🔓 [HF Models] inclusionAI - Sing-Guard-4b
https://huggingface.co/inclusionAI/Sing-Guard-4b
huggingface.co
inclusionAI/SingGuard-2b · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
🆕 [HF Models] jdopensource - JoyAI-Image-Edit-Plus-Diffusers
https://huggingface.co/jdopensource/JoyAI-Image-Edit-Plus-Diffusers
https://huggingface.co/jdopensource/JoyAI-Image-Edit-Plus-Diffusers
huggingface.co
jdopensource/JoyAI-Image-Edit-Plus-Diffusers · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
📰 HuggingFace - PP-OCRv6 on Hugging Face: 50-Language OCR from 1.5M to 34.5M Parameters
https://huggingface.co/blog/PaddlePaddle/pp-ocrv6
https://huggingface.co/blog/PaddlePaddle/pp-ocrv6
huggingface.co
PP-OCRv6 on Hugging Face: 50-Language OCR from 1.5M to 34.5M Parameters
A Blog post by PaddlePaddle on Hugging Face
🆕 [HF Models] inclusionAI - Sing-Guard-2b-GGUF
https://huggingface.co/inclusionAI/Sing-Guard-2b-GGUF
🆕 [HF Models] inclusionAI - Sing-Guard-4b-GGUF
https://huggingface.co/inclusionAI/Sing-Guard-4b-GGUF
🆕 [HF Models] inclusionAI - Sing-Guard-8b-GGUF
https://huggingface.co/inclusionAI/Sing-Guard-8b-GGUF
https://huggingface.co/inclusionAI/Sing-Guard-2b-GGUF
🆕 [HF Models] inclusionAI - Sing-Guard-4b-GGUF
https://huggingface.co/inclusionAI/Sing-Guard-4b-GGUF
🆕 [HF Models] inclusionAI - Sing-Guard-8b-GGUF
https://huggingface.co/inclusionAI/Sing-Guard-8b-GGUF
huggingface.co
inclusionAI/SingGuard-2b-GGUF · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
📰 Google DeepMind - Google DeepMind and A24 announce first-of-its-kind research partnership
https://blog.google/innovation-and-ai/models-and-research/google-deepmind/deepmind-a24-research-partnership/
https://blog.google/innovation-and-ai/models-and-research/google-deepmind/deepmind-a24-research-partnership/
Google
Google DeepMind and A24 announce first-of-its-kind research partnership
Today, Google DeepMind and A24 are announcing a first-of-its-kind partnership focused on research. The collaboration pairs a world-leading research lab with the industry…
📰 Google AI Blog - Build Cross-Language Multi-Agent Team with Google’s Agent Development Kit and A2A
How a Python agent and a Go agent collaborate on contract compliance using the Agent2Agent protocolY...
https://developers.googleblog.com/en/build-cross-language-multi-agent-team-with-google-agent-development-kit-and-a2a/
How a Python agent and a Go agent collaborate on contract compliance using the Agent2Agent protocolY...
https://developers.googleblog.com/en/build-cross-language-multi-agent-team-with-google-agent-development-kit-and-a2a/
Googleblog
Google for Developers Blog - News about Web, Mobile, AI and Cloud
How a Python agent and a Go agent collaborate on contract compliance using the Agent2Agent protocolY...
📰 NVIDIA - Enable Real-Time AI for High-Speed Data Acquisition with DAQIRI
When AlphaFold2 revolutionized drug discovery in 2020, its success relied entirely on the roughly 170,000 protein structures collected by scientists since 1971…
https://developer.nvidia.com/blog/enable-real-time-ai-for-high-speed-data-acquisition-with-daqiri/
📰 NVIDIA - Inside NVIDIA Halos for Robotics: A Full-Stack Functional Safety System for Physical AI
Physical AI—robots working autonomously alongside people in factories, warehouses, hospitals, and homes—is arriving faster than most expected.
https://developer.nvidia.com/blog/inside-nvidia-halos-for-robotics-a-full-stack-functional-safety-system-for-physical-ai/
When AlphaFold2 revolutionized drug discovery in 2020, its success relied entirely on the roughly 170,000 protein structures collected by scientists since 1971…
https://developer.nvidia.com/blog/enable-real-time-ai-for-high-speed-data-acquisition-with-daqiri/
📰 NVIDIA - Inside NVIDIA Halos for Robotics: A Full-Stack Functional Safety System for Physical AI
Physical AI—robots working autonomously alongside people in factories, warehouses, hospitals, and homes—is arriving faster than most expected.
https://developer.nvidia.com/blog/inside-nvidia-halos-for-robotics-a-full-stack-functional-safety-system-for-physical-ai/
NVIDIA Technical Blog
Enable Real-Time AI for High-Speed Data Acquisition with DAQIRI
When AlphaFold2 revolutionized drug discovery in 2020, its success relied entirely on the roughly 170,000 protein structures collected by scientists since 1971 and preserved in the Protein Data Bank.
📰 HuggingFace - Shipping huggingface_hub every week with AI, open tools, and a human in the loop
https://huggingface.co/blog/huggingface-hub-release-ci
🔓 HuggingFace - We got local models to triage the OpenClaw repo for FREE!*
https://huggingface.co/blog/local-models-pr-triage
https://huggingface.co/blog/huggingface-hub-release-ci
🔓 HuggingFace - We got local models to triage the OpenClaw repo for FREE!*
https://huggingface.co/blog/local-models-pr-triage
huggingface.co
Shipping huggingface_hub every week with AI, open tools, and a human in the loop
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
📰 HuggingFace - Build real agentic apps using CUGA: two dozen working examples on a lightweight harness
https://huggingface.co/blog/ibm-research/cuga-apps
https://huggingface.co/blog/ibm-research/cuga-apps
huggingface.co
Build real agentic apps using CUGA: two dozen working examples on a lightweight harness
A Blog post by IBM Research on Hugging Face
📰 PyTorch - Serving DeepSeek-V4 on GB300 with SGLang: 5x Higher Throughput at the Same Interactivity Since Day-0
TL;DR: DeepSeek-V4 support was live in SGLang on Day-0, but the Day-0 stack was only the starting point. Since launch, we have coordinated a set of kernel, runtime, and hardening...
https://pytorch.org/blog/serving-deepseek-v4-on-gb300-with-sglang-5x-higher-throughput-at-the-same-interactivity-since-day-0/
TL;DR: DeepSeek-V4 support was live in SGLang on Day-0, but the Day-0 stack was only the starting point. Since launch, we have coordinated a set of kernel, runtime, and hardening...
https://pytorch.org/blog/serving-deepseek-v4-on-gb300-with-sglang-5x-higher-throughput-at-the-same-interactivity-since-day-0/
🔓 HuggingFace - Experimenting with the proposed Cross-Origin Storage API in Transformers.js
https://huggingface.co/blog/cross-origin-storage
https://huggingface.co/blog/cross-origin-storage
huggingface.co
Experimenting with the proposed Cross-Origin Storage API in Transformers.js
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
📰 NVIDIA - Boost Inference Performance up to 15x on NVIDIA Blackwell Using DFlash Speculative Decoding
As AI systems move from single-turn interactions to coordinated multiagent workflows, low-latency inference becomes increasingly important.
https://developer.nvidia.com/blog/boost-inference-performance-up-to-15x-on-nvidia-blackwell-using-dflash-speculative-decoding/
📰 NVIDIA - How Telcos Build Autonomous Networks with Agentic AI
Telecom operators are adopting AI across network operations, customer care, and back-office workflows, but most are still early in the journey to autonomy.
https://developer.nvidia.com/blog/how-telcos-build-autonomous-networks-with-agentic-ai/
As AI systems move from single-turn interactions to coordinated multiagent workflows, low-latency inference becomes increasingly important.
https://developer.nvidia.com/blog/boost-inference-performance-up-to-15x-on-nvidia-blackwell-using-dflash-speculative-decoding/
📰 NVIDIA - How Telcos Build Autonomous Networks with Agentic AI
Telecom operators are adopting AI across network operations, customer care, and back-office workflows, but most are still early in the journey to autonomy.
https://developer.nvidia.com/blog/how-telcos-build-autonomous-networks-with-agentic-ai/
NVIDIA Technical Blog
Boost Inference Performance up to 15x on NVIDIA Blackwell Using DFlash Speculative Decoding
As AI systems move from single-turn interactions to coordinated multiagent workflows, low-latency inference becomes increasingly important. Autoregressive LLMs generate tokens sequentially…