📰 NVIDIA - NVIDIA Vera Storage Benchmarks: Faster Encryption, Compression, Integrity Checking, and Recovery for AI-Native Storage
Storage is an active part of every agentic AI workflow. As agents retrieve enterprise knowledge, access persistent memory, reuse key-value (KV) cache data…
https://developer.nvidia.com/blog/nvidia-vera-storage-benchmarks-faster-encryption-compression-integrity-checking-and-recovery-for-ai-native-storage/
Storage is an active part of every agentic AI workflow. As agents retrieve enterprise knowledge, access persistent memory, reuse key-value (KV) cache data…
https://developer.nvidia.com/blog/nvidia-vera-storage-benchmarks-faster-encryption-compression-integrity-checking-and-recovery-for-ai-native-storage/
NVIDIA Technical Blog
NVIDIA Vera Storage Benchmarks: Faster Encryption, Compression, Integrity Checking, and Recovery for AI-Native Storage
Storage is an active part of every agentic AI workflow. As agents retrieve enterprise knowledge, access persistent memory, reuse key-value (KV) cache data, execute tools, and generate new results…
📰 HuggingFace - Deploy local agents everywhere with LFM2.5-2.6B
https://huggingface.co/blog/LiquidAI/lfm2-5-2-6b
https://huggingface.co/blog/LiquidAI/lfm2-5-2-6b
huggingface.co
Deploy local agents everywhere with LFM2.5-2.6B
A Blog post by Liquid AI on Hugging Face
🆕 [HF Models] LiquidAI - LFM2.5-2.6B-GGUF
https://huggingface.co/LiquidAI/LFM2.5-2.6B-GGUF
🆕 [HF Models] LiquidAI - LFM2.5-2.6B-Base
https://huggingface.co/LiquidAI/LFM2.5-2.6B-Base
🆕 [HF Models] LiquidAI - LFM2.5-2.6B
https://huggingface.co/LiquidAI/LFM2.5-2.6B
https://huggingface.co/LiquidAI/LFM2.5-2.6B-GGUF
🆕 [HF Models] LiquidAI - LFM2.5-2.6B-Base
https://huggingface.co/LiquidAI/LFM2.5-2.6B-Base
🆕 [HF Models] LiquidAI - LFM2.5-2.6B
https://huggingface.co/LiquidAI/LFM2.5-2.6B
huggingface.co
LiquidAI/LFM2.5-2.6B-Base · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
📰 Mistral - Introducing Shieldstral.
Shieldstral introduces a 3B open-weights multimodal safety classifier that outperforms models up to 7x its size.
https://mistral.ai/news/shieldstral/
Shieldstral introduces a 3B open-weights multimodal safety classifier that outperforms models up to 7x its size.
https://mistral.ai/news/shieldstral/
Mistral AI
Introducing Shieldstral. | Mistral AI
Shieldstral introduces a 3B open-weights multimodal safety classifier that outperforms models up to 7x its size.
📰 Google DeepMind - The latest AI news we announced in July 2026
https://blog.google/innovation-and-ai/technology/ai/google-ai-updates-july-2026/
https://blog.google/innovation-and-ai/technology/ai/google-ai-updates-july-2026/
Google
The latest AI news we announced in July 2026
Here are Google’s latest AI updates from July 2026
📰 Google AI Blog - A unified API for AI model routing
Google Cloud API Gateway now offers a model routing feature in Public Preview, allowing developers to dynamically route traffic to models like Gemini, Claude, or OpenAI OSS-GPT without hardcoding endpoints or managing open-source proxies. Developers can easily configure these routing rules directly within their OpenAPI 3.x specifications by mapping virtual model names to specific backend targets on a shared host. Once deployed, the Gateway acts as a serverless ingress layer that accepts standard OpenAI-compatible requests, automatically transcodes the payload to the native schema of the target model, and routes the traffic on the fly.
https://developers.googleblog.com/en/a-unified-api-for-ai-model-routing/
Google Cloud API Gateway now offers a model routing feature in Public Preview, allowing developers to dynamically route traffic to models like Gemini, Claude, or OpenAI OSS-GPT without hardcoding endpoints or managing open-source proxies. Developers can easily configure these routing rules directly within their OpenAPI 3.x specifications by mapping virtual model names to specific backend targets on a shared host. Once deployed, the Gateway acts as a serverless ingress layer that accepts standard OpenAI-compatible requests, automatically transcodes the payload to the native schema of the target model, and routes the traffic on the fly.
https://developers.googleblog.com/en/a-unified-api-for-ai-model-routing/
Googleblog
Google for Developers Blog - News about Web, Mobile, AI and Cloud
Discover how developers can configure Google Cloud API Gateway to dynamically route OpenAI-compatible requests without managing open-source proxies.
📰 LMSys - SpecForge v0.3.0: a Unified Disaggregated and Colocated Speculative Decoding Stack, and New Open SpecBundle Draft Models
https://lmsys.org/blog/2026-08-04-specforge-v0-3
https://lmsys.org/blog/2026-08-04-specforge-v0-3
www.lmsys.org
SpecForge v0.3.0: a Unified Disaggregated and Colocated Speculative Decoding Stack, and New Open SpecBundle Draft Models
When we first released SpecForge, a training job owned both the frozen target model and the draft model being optimized. This made EAGLE3 draft-model training practical and directly compatible with SG...
📰 Anthropic - Mariano-Florentino (Tino) Cuéllar to join Anthropic as Chief Global Affairs Officer
https://www.anthropic.com/news/tino-cuellar
https://www.anthropic.com/news/tino-cuellar
Anthropic
Mariano-Florentino (Tino) Cuéllar to join Anthropic as Chief Global Affairs Officer
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
📰 OpenAI - New ways to learn and teach with ChatGPT Work and Codex
Explore new education plugins for ChatGPT Work and Codex that help K–12 teachers, college educators, and students learn, teach, research, and build.
https://openai.com/index/learn-teach-chatgpt-work-codex
📰 OpenAI - Apple is getting this wrong
OpenAI addresses Apple’s baseless lawsuit, corrects claims about its employees, and shares messages documenting what happened.
https://openai.com/index/apple-is-getting-this-wrong
📰 OpenAI - How we built a realtime system for responsive voice AI in six months
GPT-Live enables continuous voice interaction with AI, using a turnless speech model and low-latency architecture for faster, more natural conversations.
https://openai.com/index/continuous-voice-interaction-with-gpt-live
Explore new education plugins for ChatGPT Work and Codex that help K–12 teachers, college educators, and students learn, teach, research, and build.
https://openai.com/index/learn-teach-chatgpt-work-codex
📰 OpenAI - Apple is getting this wrong
OpenAI addresses Apple’s baseless lawsuit, corrects claims about its employees, and shares messages documenting what happened.
https://openai.com/index/apple-is-getting-this-wrong
📰 OpenAI - How we built a realtime system for responsive voice AI in six months
GPT-Live enables continuous voice interaction with AI, using a turnless speech model and low-latency architecture for faster, more natural conversations.
https://openai.com/index/continuous-voice-interaction-with-gpt-live
OpenAI
New ways to learn and teach with ChatGPT Work and Codex
Explore new education plugins for ChatGPT Work and Codex that help K–12 teachers, college educators, and students learn, teach, research, and build.
📰 NVIDIA - Beyond VLAs: How World Action Models Reshape Robot Manipulation
A central challenge in robotics is building policies that generalize beyond the demonstrations they’re trained on. A policy that succeeds in a training scene…
https://developer.nvidia.com/blog/beyond-vlas-how-world-action-models-reshape-robot-manipulation/
A central challenge in robotics is building policies that generalize beyond the demonstrations they’re trained on. A policy that succeeds in a training scene…
https://developer.nvidia.com/blog/beyond-vlas-how-world-action-models-reshape-robot-manipulation/
NVIDIA Technical Blog
Beyond VLAs: How World Action Models Reshape Robot Manipulation
A central challenge in robotics is building policies that generalize beyond the demonstrations they’re trained on. A policy that succeeds in a training scene often fails when object shapes, positions…
📰 PyTorch - PyTorch by the Sea: The inaugural Santa Cruz PyTorch Meetup
TL;DR The inaugural Santa Cruz PyTorch Meetup brought together 45 local engineers, students, and leaders for GPU/CUDA talks and lightning presentations on chemistry, plant health, and autonomous driving – demonstrating...
https://pytorch.org/blog/pytorch-by-the-sea-the-inaugural-santa-cruz-pytorch-meetup/
TL;DR The inaugural Santa Cruz PyTorch Meetup brought together 45 local engineers, students, and leaders for GPU/CUDA talks and lightning presentations on chemistry, plant health, and autonomous driving – demonstrating...
https://pytorch.org/blog/pytorch-by-the-sea-the-inaugural-santa-cruz-pytorch-meetup/