π° HuggingFace - Security incident disclosure β July 2026
https://huggingface.co/blog/security-incident-july-2026
https://huggingface.co/blog/security-incident-july-2026
huggingface.co
Security incident disclosure β July 2026
Weβre on a journey to advance and democratize artificial intelligence through open source and open science.
π° HuggingFace - Newer Models, Same Advantage
https://huggingface.co/blog/Dharma-AI/newer-models-same-advantages
https://huggingface.co/blog/Dharma-AI/newer-models-same-advantages
huggingface.co
Newer Models, Same Advantage
A Blog post by Dharma-AI on Hugging Face
π° HuggingFace - NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval
https://huggingface.co/blog/nvidia/nemotron-3-embed-wins-rteb
https://huggingface.co/blog/nvidia/nemotron-3-embed-wins-rteb
huggingface.co
NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval
A Blog post by NVIDIA on Hugging Face
π° Google AI Blog - Expanding Choice in Gemini Enterprise Agent Platform: Introducing Grounding with Parallel Web Search
Google Cloud has partnered with Parallel Web Systems to natively integrate Parallel's search infrastructure as a web grounding provider on the Gemini Enterprise Agent Platform. This integration enables developers to anchor their AI agents in verifiable, real-time web results, significantly improving factual accuracy for complex enterprise workflows. Additionally, the partnership offers expanded architectural flexibility, allowing users to programmatically extract, permanently cache, and process web data alongside other large language models.
https://developers.googleblog.com/en/expanding-choice-in-gemini-enterprise-agent-platform-introducing-grounding-with-parallel-web-search/
Google Cloud has partnered with Parallel Web Systems to natively integrate Parallel's search infrastructure as a web grounding provider on the Gemini Enterprise Agent Platform. This integration enables developers to anchor their AI agents in verifiable, real-time web results, significantly improving factual accuracy for complex enterprise workflows. Additionally, the partnership offers expanded architectural flexibility, allowing users to programmatically extract, permanently cache, and process web data alongside other large language models.
https://developers.googleblog.com/en/expanding-choice-in-gemini-enterprise-agent-platform-introducing-grounding-with-parallel-web-search/
Googleblog
Google for Developers Blog - News about Web, Mobile, AI and Cloud
Ground Gemini models with real-time web data using Parallel Web Systems. Extract, cache, and post-process search results for complex multi-agent architectures.
π° OpenAI - Why teens deserve access to safe AI
https://openai.com/index/why-teens-deserve-access-safe-ai
π° OpenAI - How Cars24 scales conversations and builds faster with OpenAI
https://openai.com/index/cars24
https://openai.com/index/why-teens-deserve-access-safe-ai
π° OpenAI - How Cars24 scales conversations and builds faster with OpenAI
https://openai.com/index/cars24
OpenAI
Why teens deserve access to safe AI
Learn how OpenAI is making ChatGPT safer for teens with age-appropriate protections, learning tools, parental controls, and expert partnerships.
π° NVIDIA - Integrating Context-Aware Video AI Agents Into Enterprise Workflows
A video analytics AI agent that can perceive, reason, and act based on massive amounts of video footage must be integrated with existing workflows andβ¦
https://developer.nvidia.com/blog/integrating-context-aware-video-ai-agents-into-enterprise-workflows/
π° NVIDIA - Scaling Agentic AI Factories Through Extreme Co-Design with NVIDIA BlueField
Agentic AI changes the infrastructure pattern for AI factories. One request can trigger many model calls, tool calls, memory lookups, policy checksβ¦
https://developer.nvidia.com/blog/scaling-agentic-ai-factories-through-extreme-co-design-with-nvidia-bluefield/
π° NVIDIA - Build a Multi-Camera 3D Tracking Application with NVIDIA DeepStream 9.1 Skills
Developers building video analytics applications across large spaces must track the same object as it moves between camera views. Single-camera 2D trackingβ¦
https://developer.nvidia.com/blog/build-a-multi-camera-3d-tracking-application-with-nvidia-deepstream-9-1-skills/
A video analytics AI agent that can perceive, reason, and act based on massive amounts of video footage must be integrated with existing workflows andβ¦
https://developer.nvidia.com/blog/integrating-context-aware-video-ai-agents-into-enterprise-workflows/
π° NVIDIA - Scaling Agentic AI Factories Through Extreme Co-Design with NVIDIA BlueField
Agentic AI changes the infrastructure pattern for AI factories. One request can trigger many model calls, tool calls, memory lookups, policy checksβ¦
https://developer.nvidia.com/blog/scaling-agentic-ai-factories-through-extreme-co-design-with-nvidia-bluefield/
π° NVIDIA - Build a Multi-Camera 3D Tracking Application with NVIDIA DeepStream 9.1 Skills
Developers building video analytics applications across large spaces must track the same object as it moves between camera views. Single-camera 2D trackingβ¦
https://developer.nvidia.com/blog/build-a-multi-camera-3d-tracking-application-with-nvidia-deepstream-9-1-skills/
NVIDIA Technical Blog
Integrating Context-Aware Video AI Agents Into Enterprise Workflows
A video analytics AI agent that can perceive, reason, and act based on massive amounts of video footage must be integrated with existing workflows and applications to be useful.
π° HuggingFace - Fine-tune video and image models at scale with NVIDIA NeMo Automodel and π€ Diffusers
https://huggingface.co/blog/nvidia/scale-diffusers-finetuning-nemo-automodel
https://huggingface.co/blog/nvidia/scale-diffusers-finetuning-nemo-automodel
huggingface.co
Fine-tune video and image models at scale with NVIDIA NeMo Automodel and π€ Diffusers
A Blog post by NVIDIA on Hugging Face
π Google AI Blog - Evolving Spec-Driven Development: Conductor Now Supports Antigravity
Conductor has evolved from a Gemini CLI extension into a portable plugin, bringing conversational Spec-Driven Development (SDD) to ecosystems like Antigravity CLI and Claude. Rather than relying on strict command sequences, developers can now chat naturally with their AI assistant while it dynamically manages persistent markdown artifacts (like spec.md and plan.md) in the background. This update eliminates workflow friction while ensuring your repository remains a version-controlled, single source of truth for your project's architecture and state across different AI tools.
https://developers.googleblog.com/en/evolving-spec-driven-development-conductor-now-supports-antigravity/
Conductor has evolved from a Gemini CLI extension into a portable plugin, bringing conversational Spec-Driven Development (SDD) to ecosystems like Antigravity CLI and Claude. Rather than relying on strict command sequences, developers can now chat naturally with their AI assistant while it dynamically manages persistent markdown artifacts (like spec.md and plan.md) in the background. This update eliminates workflow friction while ensuring your repository remains a version-controlled, single source of truth for your project's architecture and state across different AI tools.
https://developers.googleblog.com/en/evolving-spec-driven-development-conductor-now-supports-antigravity/
Googleblog
Google for Developers Blog - News about Web, Mobile, AI and Cloud
Upgrade to the Conductor Plugin for conversational Spec-Driven Development. Chat naturally with AI to manage markdown specs across tools like Antigravity CLI.
π xAI - Grok Build is Now Open Source
https://x.ai//news/grok-build-open-source
π xAI - Automations in Grok
https://x.ai//news/grok-automations
https://x.ai//news/grok-build-open-source
π xAI - Automations in Grok
https://x.ai//news/grok-automations
x.ai
Grok Build is Now Open Source
Explore the harness behind our coding agent and TUI.
π° Claude Blog - Working at the frontier: How Cursor knew Claude Fable 5 was ready for the hardest 1% of problems
https://claude.com/blog/working-at-the-frontier-cursor
π° Claude Blog - Zero risk isn't the job: a CISO's guide to agentic AI
https://claude.com/blog/ciso-guide-to-agentic-ai
π° Claude Blog - How Anthropic runs large-scale code migrations with Claude Code
https://claude.com/blog/ai-code-migration
π° Claude Blog - Working with Claude Fable 5 in Claude Cowork
https://claude.com/blog/working-with-claude-fable-5-in-claude-cowork
π° Claude Blog - Working at the frontier: Why Base44 trusts Claude Fable 5 with their most challenging engineering work
https://claude.com/blog/working-at-the-frontier-why-base44-trusts-claude-fable-5-with-their-most-challenging-engineering-work
https://claude.com/blog/working-at-the-frontier-cursor
π° Claude Blog - Zero risk isn't the job: a CISO's guide to agentic AI
https://claude.com/blog/ciso-guide-to-agentic-ai
π° Claude Blog - How Anthropic runs large-scale code migrations with Claude Code
https://claude.com/blog/ai-code-migration
π° Claude Blog - Working with Claude Fable 5 in Claude Cowork
https://claude.com/blog/working-with-claude-fable-5-in-claude-cowork
π° Claude Blog - Working at the frontier: Why Base44 trusts Claude Fable 5 with their most challenging engineering work
https://claude.com/blog/working-at-the-frontier-why-base44-trusts-claude-fable-5-with-their-most-challenging-engineering-work
Claude
How Cursor knew Claude Fable 5 was ready for the hardest 1% of problems | Claude by Anthropic
How Anthropic's Claude Fable 5 beat CursorBench and expanded what's possible for Cursor and agentic coding.
π [GitHub Releases] turboderp-org/exllamav3 - 1.1.0
https://github.com/turboderp-org/exllamav3/releases/tag/v1.1.0
https://github.com/turboderp-org/exllamav3/releases/tag/v1.1.0
GitHub
Release 1.1.0 Β· turboderp-org/exllamav3
An optimized quantization and inference library for running LLMs locally on modern consumer-class GPUs - Release 1.1.0 Β· turboderp-org/exllamav3
π [HF Models] openbmb - MiniCPM-RobotTrack
https://huggingface.co/openbmb/MiniCPM-RobotTrack
π [HF Models] openbmb - MiniCPM-RobotManip
https://huggingface.co/openbmb/MiniCPM-RobotManip
https://huggingface.co/openbmb/MiniCPM-RobotTrack
π [HF Models] openbmb - MiniCPM-RobotManip
https://huggingface.co/openbmb/MiniCPM-RobotManip
huggingface.co
openbmb/MiniCPM-RobotTrack Β· Hugging Face
Weβre on a journey to advance and democratize artificial intelligence through open source and open science.
π [GitHub Releases] PygmalionAI/aphrodite-engine - v0.22.0
https://github.com/dphnAI/sonar/releases/tag/v0.22.0
https://github.com/dphnAI/sonar/releases/tag/v0.22.0
GitHub
Release v0.22.0 Β· dphnAI/sonar
What's Changed
feat: add native Metal support by @AlpinDale in #1668
chore: optimize metal backend performance by @AlpinDale in #1669
perf: optimize GDN performance on Metal by @AlpinDale in #...
feat: add native Metal support by @AlpinDale in #1668
chore: optimize metal backend performance by @AlpinDale in #1669
perf: optimize GDN performance on Metal by @AlpinDale in #...
ποΈ Weekly GitHub Activity
π¦ llama.cpp
β Release: b9966 β b10068
β 102 commits
- Added support for Hunyuan 3 (hy_v3) with MTP speculative decoding #25395 #25641
- Added support for Minimax2 Eagle3 speculative decoding 259ae1d
- Added support for BitNetForCausalLM GGUF conversion #25769
- Implemented GGML_OP_LIGHTNING_INDEXER for DeepSeek V3.2/V4 on CPU and CUDA #24231 #25545
- Added fused hyper-connection ops for DeepSeek V4 to reduce graph splits #25585 #25702
- Added CUDA Virtual Devices support and enabled CUDA graphs on Volta and Turing architectures #25228 #25749
- Added Flash Attention via oneDNN graph API for SYCL on Intel Battlemage #25222
- Optimized CUDA MoE gate/up activation quantization, improving prefill times on RTX 5090 and Blackwell #25441
- Added auto-download of DeepSeek-Flash and Eagle3 speculative decoding sidecars from Hugging Face #25811
- Server now supports CORS configuration options and accepts null sampling parameters to request defaults #25655 #25538
- Added KleidiAI SME2 f32 kernel and improved hardware-specific kernel dispatch #24414 #25478
- Fixed CUDA crash when querying memory on devices with no available memory #25157
- Fixed Tensor Parallel execution for Phi3, Bert, Plamo2/3, and ChatGLM #25536
- Fixed quantization crash on DeepSeek-V4 i32 routing tables #25787
π All changes | Latest release
π¨ stable-diffusion.cpp
β Release: master-775-b5d8120 β master-782-b290693
β 7 commits
- Support for AnimateDiff SD 1.5 motion modules v2 and v3 #1784 with img2video capabilities via the --init-img parameter #1789
- Support for ADetailer #1785
- Support for PiD 1.5 #1790
- Configurable reference image processing for edit models #1780
- Fixed cross attention and output projection token protection for Anima LoRAs #1786
π All changes | Latest release
π€ Fresh models trending on HuggingFace:
thinkingmachines/Inkling β‘1060
OpenMOSS-Team/MOSS-VL-Realtime β‘76
ai-sage/GigaAM-Multilingual β‘56
nineninesix/diamond-1.0 β‘43
ai-sage/GigaChat3.1-Audio-10B-A1.8B β‘38
rzgar/Bernini-R-S2V β‘37
acvlab/ABot-World-0-5B-LF β‘29
fal/ideogram-v4-instant β‘27
InternScience/Agents-A1-4B β‘26
sensenova/SenseNova-U1-8B-MoT-Infographic-V3 β‘26
OpenMOSS-Team/MOSS-VL-Instruct-0708 β‘23
fal/ideogram-v4-fast β‘23
t-tech/T-Search β‘23
GnLOLot/MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking β‘23
mente-ai/uyu-2-28B β‘20
OpenMOSS-Team/MOSS-VL-Base-0708 β‘17
yijunwang2/krea2-outpaint β‘17
yijunwang2/krea2-reid β‘15
π¦ llama.cpp
β Release: b9966 β b10068
β 102 commits
- Added support for Hunyuan 3 (hy_v3) with MTP speculative decoding #25395 #25641
- Added support for Minimax2 Eagle3 speculative decoding 259ae1d
- Added support for BitNetForCausalLM GGUF conversion #25769
- Implemented GGML_OP_LIGHTNING_INDEXER for DeepSeek V3.2/V4 on CPU and CUDA #24231 #25545
- Added fused hyper-connection ops for DeepSeek V4 to reduce graph splits #25585 #25702
- Added CUDA Virtual Devices support and enabled CUDA graphs on Volta and Turing architectures #25228 #25749
- Added Flash Attention via oneDNN graph API for SYCL on Intel Battlemage #25222
- Optimized CUDA MoE gate/up activation quantization, improving prefill times on RTX 5090 and Blackwell #25441
- Added auto-download of DeepSeek-Flash and Eagle3 speculative decoding sidecars from Hugging Face #25811
- Server now supports CORS configuration options and accepts null sampling parameters to request defaults #25655 #25538
- Added KleidiAI SME2 f32 kernel and improved hardware-specific kernel dispatch #24414 #25478
- Fixed CUDA crash when querying memory on devices with no available memory #25157
- Fixed Tensor Parallel execution for Phi3, Bert, Plamo2/3, and ChatGLM #25536
- Fixed quantization crash on DeepSeek-V4 i32 routing tables #25787
π All changes | Latest release
π¨ stable-diffusion.cpp
β Release: master-775-b5d8120 β master-782-b290693
β 7 commits
- Support for AnimateDiff SD 1.5 motion modules v2 and v3 #1784 with img2video capabilities via the --init-img parameter #1789
- Support for ADetailer #1785
- Support for PiD 1.5 #1790
- Configurable reference image processing for edit models #1780
- Fixed cross attention and output projection token protection for Anima LoRAs #1786
π All changes | Latest release
π€ Fresh models trending on HuggingFace:
thinkingmachines/Inkling β‘1060
OpenMOSS-Team/MOSS-VL-Realtime β‘76
ai-sage/GigaAM-Multilingual β‘56
nineninesix/diamond-1.0 β‘43
ai-sage/GigaChat3.1-Audio-10B-A1.8B β‘38
rzgar/Bernini-R-S2V β‘37
acvlab/ABot-World-0-5B-LF β‘29
fal/ideogram-v4-instant β‘27
InternScience/Agents-A1-4B β‘26
sensenova/SenseNova-U1-8B-MoT-Infographic-V3 β‘26
OpenMOSS-Team/MOSS-VL-Instruct-0708 β‘23
fal/ideogram-v4-fast β‘23
t-tech/T-Search β‘23
GnLOLot/MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking β‘23
mente-ai/uyu-2-28B β‘20
OpenMOSS-Team/MOSS-VL-Base-0708 β‘17
yijunwang2/krea2-outpaint β‘17
yijunwang2/krea2-reid β‘15
GitHub
model: add Hy3 (hy_v3) support with MTP speculative decoding by satindergrewal Β· Pull Request #25395 Β· ggml-org/llama.cpp
Overview
Adds support for Tencent's Hy3 (hy_v3 / HYV3ForCausalLM, 299B MoE, 80 layers + 1 MTP layer), including its multi-token-prediction head as a draft-mtp speculative target. Addresses ...
Adds support for Tencent's Hy3 (hy_v3 / HYV3ForCausalLM, 299B MoE, 80 layers + 1 MTP layer), including its multi-token-prediction head as a draft-mtp speculative target. Addresses ...
π [HF Models] nvidia - Cosmos3-Edge
https://huggingface.co/nvidia/Cosmos3-Edge
π [HF Models] nvidia - Cosmos3-Super-Text2Image-4Step
https://huggingface.co/nvidia/Cosmos3-Super-Text2Image-4Step
π [HF Models] nvidia - Cosmos3-Super-Image2Video-4Step
https://huggingface.co/nvidia/Cosmos3-Super-Image2Video-4Step
π [HF Models] nvidia - Cosmos3-Edge-Policy-DROID
https://huggingface.co/nvidia/Cosmos3-Edge-Policy-DROID
https://huggingface.co/nvidia/Cosmos3-Edge
π [HF Models] nvidia - Cosmos3-Super-Text2Image-4Step
https://huggingface.co/nvidia/Cosmos3-Super-Text2Image-4Step
π [HF Models] nvidia - Cosmos3-Super-Image2Video-4Step
https://huggingface.co/nvidia/Cosmos3-Super-Image2Video-4Step
π [HF Models] nvidia - Cosmos3-Edge-Policy-DROID
https://huggingface.co/nvidia/Cosmos3-Edge-Policy-DROID
huggingface.co
nvidia/Cosmos3-Edge Β· Hugging Face
Weβre on a journey to advance and democratize artificial intelligence through open source and open science.