📰 NVIDIA - Create a LangChain Deep Agents Harness Profile for NVIDIA Nemotron 3 Ultra to Improve Performance
Agentic systems often face a trade-off between accuracy and cost. The highest-performing proprietary frontier models and harnesses provide top accuracy but are…
https://developer.nvidia.com/blog/create-a-langchain-deep-agents-harness-profile-for-nvidia-nemotron-3-ultra-to-improve-performance/
Agentic systems often face a trade-off between accuracy and cost. The highest-performing proprietary frontier models and harnesses provide top accuracy but are…
https://developer.nvidia.com/blog/create-a-langchain-deep-agents-harness-profile-for-nvidia-nemotron-3-ultra-to-improve-performance/
NVIDIA Technical Blog
Create a LangChain Deep Agents Harness Profile for NVIDIA Nemotron 3 Ultra to Improve Performance
Agentic systems often face a trade-off between accuracy and cost. The highest-performing proprietary frontier models and harnesses provide top accuracy but are expensive. Fine-tuning offers one way to…
📰 OpenAI - Our approach to government and national security partnerships
https://openai.com/index/government-national-security-partnerships
https://openai.com/index/government-national-security-partnerships
OpenAI
Our approach to government and national security partnerships
Learn how OpenAI approaches government and national security partnerships, with principles for responsible AI use, democratic accountability, and public safety.
📰 Meta AI - Introducing Muse Spark 1.1
https://ai.meta.com/blog/introducing-muse-spark-meta-model-api/
https://ai.meta.com/blog/introducing-muse-spark-meta-model-api/
Meta AI
Introducing Muse Spark 1.1
📰 Mistral - Your Prompts and Skills need a system of record.
Studio provides a system of record for AI prompts and skills—versioned, owned, and traceable. Iterate fast, ship with control, and ensure consistent AI behavior.
https://mistral.ai/news/manage-prompts-and-skills-in-studio/
Studio provides a system of record for AI prompts and skills—versioned, owned, and traceable. Iterate fast, ship with control, and ensure consistent AI behavior.
https://mistral.ai/news/manage-prompts-and-skills-in-studio/
Mistral AI
Version control for prompts & skills in Studio | Mistral
Studio provides a system of record for AI prompts and skills—versioned, owned, and traceable. Iterate fast, ship with control, and ensure consistent AI behavior.
📰 Google DeepMind - We're rolling out AlphaEvolve widely to solve Google Cloud customers' hardest problems.
https://blog.google/innovation-and-ai/infrastructure-and-cloud/google-cloud/alphaevolve-on-cloud/
https://blog.google/innovation-and-ai/infrastructure-and-cloud/google-cloud/alphaevolve-on-cloud/
Google
We're rolling out AlphaEvolve widely to solve Google Cloud customers' hardest problems.
Finding the most efficient algorithm — whether designing a microchip, routing a logistics network or accelerating medical research — can be challenging, with many possib…
🔓 OpenAI - ChatGPT Sites Data Processing Addendum
https://openai.com/policies/chatgpt-sites-data-processing-addendum
🔓 OpenAI - ChatGPT Sites Terms
https://openai.com/policies/chatgpt-sites-terms
🔓 OpenAI - App Developer Terms
https://openai.com/policies/developer-apps-terms
🔓 OpenAI - Take on your most ambitious work with ChatGPT
https://openai.com/chatgpt-work
🔓 OpenAI - GPT-5.6: Frontier intelligence that scales with your ambition
https://openai.com/index/gpt-5-6
🔓 OpenAI - OpenAI Bio Bug Bounty
https://openai.com/index/bio-bug-bounty
🔓 OpenAI - ChatGPT is now a partner for your most ambitious work
https://openai.com/index/chatgpt-for-your-most-ambitious-work
https://openai.com/policies/chatgpt-sites-data-processing-addendum
🔓 OpenAI - ChatGPT Sites Terms
https://openai.com/policies/chatgpt-sites-terms
🔓 OpenAI - App Developer Terms
https://openai.com/policies/developer-apps-terms
🔓 OpenAI - Take on your most ambitious work with ChatGPT
https://openai.com/chatgpt-work
🔓 OpenAI - GPT-5.6: Frontier intelligence that scales with your ambition
https://openai.com/index/gpt-5-6
🔓 OpenAI - OpenAI Bio Bug Bounty
https://openai.com/index/bio-bug-bounty
🔓 OpenAI - ChatGPT is now a partner for your most ambitious work
https://openai.com/index/chatgpt-for-your-most-ambitious-work
OpenAI
ChatGPT Sites Data Processing Addendum
Read the ChatGPT Sites Data Processing Addendum to understand how OpenAI processes, protects, and manages personal data for published ChatGPT Sites.
📰 Claude Blog - Working at the frontier: How Thomson Reuters builds AI for high-stakes professional work
https://claude.com/blog/working-at-the-frontier-how-thomson-reuters-builds-ai-for-high--stakes-professional-work
📰 Claude Blog - How Anthropic's marketing operations team uses Claude Cowork to automate reporting and campaign builds
https://claude.com/blog/how-anthropics-marketing-operations-team-uses-claude-cowork-to-automate-reporting-and-campaign-builds
📰 Claude Blog - Bringing Claude Code and Claude Cowork to government
https://claude.com/blog/bringing-claude-code-and-claude-cowork-to-government
📰 Claude Blog - Choosing a Claude model and effort level in Claude Code
https://claude.com/blog/claude-model-and-effort-level-in-claude-code
📰 Claude Blog - Claude Cowork is coming to mobile and web
https://claude.com/blog/cowork-web-mobile
📰 Claude Blog - How people are using Claude Cowork
https://claude.com/blog/how-people-are-using-claude-cowork
📰 Claude Blog - A field guide to Claude Fable 5: Finding your unknowns
https://claude.com/blog/a-field-guide-to-claude-fable-finding-your-unknowns
https://claude.com/blog/working-at-the-frontier-how-thomson-reuters-builds-ai-for-high--stakes-professional-work
📰 Claude Blog - How Anthropic's marketing operations team uses Claude Cowork to automate reporting and campaign builds
https://claude.com/blog/how-anthropics-marketing-operations-team-uses-claude-cowork-to-automate-reporting-and-campaign-builds
📰 Claude Blog - Bringing Claude Code and Claude Cowork to government
https://claude.com/blog/bringing-claude-code-and-claude-cowork-to-government
📰 Claude Blog - Choosing a Claude model and effort level in Claude Code
https://claude.com/blog/claude-model-and-effort-level-in-claude-code
📰 Claude Blog - Claude Cowork is coming to mobile and web
https://claude.com/blog/cowork-web-mobile
📰 Claude Blog - How people are using Claude Cowork
https://claude.com/blog/how-people-are-using-claude-cowork
📰 Claude Blog - A field guide to Claude Fable 5: Finding your unknowns
https://claude.com/blog/a-field-guide-to-claude-fable-finding-your-unknowns
Claude
Working at the frontier: How Thomson Reuters builds AI for high- stakes professional work | Claude by Anthropic
Why the team at Thomson Reuters considers Claude Fable 5 a critical evolution in what’s possible with AI for knowledge work.
📰 Anthropic Research - An off switch for dual-use knowledge in AI models
https://www.anthropic.com/research/off-switch-dual-use
https://www.anthropic.com/research/off-switch-dual-use
Anthropic
An off switch for dual-use knowledge in AI models
New results on a method of controlling access to potentially dangerous AI capabilities
📰 Anthropic - Inviting hard questions
https://www.anthropic.com/news/hard-questions
📰 Anthropic - Ben Bernanke appointed to Anthropic’s Long-Term Benefit Trust
https://www.anthropic.com/news/ben-bernanke
📰 Anthropic - Introducing a way to reflect on how you use Claude
https://www.anthropic.com/news/reflect-with-claude
https://www.anthropic.com/news/hard-questions
📰 Anthropic - Ben Bernanke appointed to Anthropic’s Long-Term Benefit Trust
https://www.anthropic.com/news/ben-bernanke
📰 Anthropic - Introducing a way to reflect on how you use Claude
https://www.anthropic.com/news/reflect-with-claude
Anthropic
Inviting hard questions
We're asking the public for their hardest questions about AI, and committing to show our work as we address them.
📰 OpenAI - GPT-5.6 is now the preferred model in Microsoft 365 Copilot
https://openai.com/index/gpt-5-6-preferred-model-microsoft-365-copilot
🔓 OpenAI - GPT-5.6 System Card
https://deploymentsafety.openai.com/gpt-5-6
https://openai.com/index/gpt-5-6-preferred-model-microsoft-365-copilot
🔓 OpenAI - GPT-5.6 System Card
https://deploymentsafety.openai.com/gpt-5-6
OpenAI
GPT-5.6 is now the preferred model in Microsoft 365 Copilot
Learn how GPT-5.6 powers Microsoft 365 Copilot with stronger AI capabilities across Word, Excel, PowerPoint, Chat, and Cowork for faster, higher-quality work.
🆕 [HF Models] nvidia - Kimi-K2.7-Code-DFlash
https://huggingface.co/nvidia/Kimi-K2.7-Code-DFlash
🔓 [HF Models] nvidia - ARDY-Core-RP-20FPS-Horizon40
https://huggingface.co/nvidia/ARDY-Core-RP-20FPS-Horizon40
🔓 [HF Models] nvidia - ARDY-G1-RP-25FPS-Horizon52
https://huggingface.co/nvidia/ARDY-G1-RP-25FPS-Horizon52
🔓 [HF Models] nvidia - ARDY-Core-RP-20FPS-Horizon8
https://huggingface.co/nvidia/ARDY-Core-RP-20FPS-Horizon8
🔓 [HF Models] nvidia - ARDY-G1-RP-25FPS-Horizon8
https://huggingface.co/nvidia/ARDY-G1-RP-25FPS-Horizon8
https://huggingface.co/nvidia/Kimi-K2.7-Code-DFlash
🔓 [HF Models] nvidia - ARDY-Core-RP-20FPS-Horizon40
https://huggingface.co/nvidia/ARDY-Core-RP-20FPS-Horizon40
🔓 [HF Models] nvidia - ARDY-G1-RP-25FPS-Horizon52
https://huggingface.co/nvidia/ARDY-G1-RP-25FPS-Horizon52
🔓 [HF Models] nvidia - ARDY-Core-RP-20FPS-Horizon8
https://huggingface.co/nvidia/ARDY-Core-RP-20FPS-Horizon8
🔓 [HF Models] nvidia - ARDY-G1-RP-25FPS-Horizon8
https://huggingface.co/nvidia/ARDY-G1-RP-25FPS-Horizon8
📰 NVIDIA - Synthetic Data Generation for Financial AI Research with NVIDIA NeMo
Fine-tuning LLMs for financial natural language processing (NLP) is constrained by limited, imbalanced data. Real-world financial news overrepresents earnings…
https://developer.nvidia.com/blog/synthetic-data-generation-for-financial-ai-research-with-nvidia-nemo/
Fine-tuning LLMs for financial natural language processing (NLP) is constrained by limited, imbalanced data. Real-world financial news overrepresents earnings…
https://developer.nvidia.com/blog/synthetic-data-generation-for-financial-ai-research-with-nvidia-nemo/
NVIDIA Technical Blog
Synthetic Data Generation for Financial AI Research with NVIDIA NeMo
Fine-tuning LLMs for financial natural language processing (NLP) is constrained by limited, imbalanced data. Real-world financial news overrepresents earnings and stock movements…
🆕 [HF Models] inclusionAI - SingGuard-0.8b-GGUF
https://huggingface.co/inclusionAI/SingGuard-0.8b-GGUF
🔓 [HF Models] inclusionAI - SingGuard-2b
https://huggingface.co/inclusionAI/SingGuard-2b
🔓 [HF Models] inclusionAI - SingGuard-8b
https://huggingface.co/inclusionAI/SingGuard-8b
🔓 [HF Models] inclusionAI - SingGuard-4b
https://huggingface.co/inclusionAI/SingGuard-4b
🆕 [HF Models] inclusionAI - SingGuard-8b-GGUF
https://huggingface.co/inclusionAI/SingGuard-8b-GGUF
🆕 [HF Models] inclusionAI - SingGuard-4b-GGUF
https://huggingface.co/inclusionAI/SingGuard-4b-GGUF
🆕 [HF Models] inclusionAI - SingGuard-2b-GGUF
https://huggingface.co/inclusionAI/SingGuard-2b-GGUF
🆕 [HF Models] inclusionAI - SingGuard-0.8b
https://huggingface.co/inclusionAI/SingGuard-0.8b
https://huggingface.co/inclusionAI/SingGuard-0.8b-GGUF
🔓 [HF Models] inclusionAI - SingGuard-2b
https://huggingface.co/inclusionAI/SingGuard-2b
🔓 [HF Models] inclusionAI - SingGuard-8b
https://huggingface.co/inclusionAI/SingGuard-8b
🔓 [HF Models] inclusionAI - SingGuard-4b
https://huggingface.co/inclusionAI/SingGuard-4b
🆕 [HF Models] inclusionAI - SingGuard-8b-GGUF
https://huggingface.co/inclusionAI/SingGuard-8b-GGUF
🆕 [HF Models] inclusionAI - SingGuard-4b-GGUF
https://huggingface.co/inclusionAI/SingGuard-4b-GGUF
🆕 [HF Models] inclusionAI - SingGuard-2b-GGUF
https://huggingface.co/inclusionAI/SingGuard-2b-GGUF
🆕 [HF Models] inclusionAI - SingGuard-0.8b
https://huggingface.co/inclusionAI/SingGuard-0.8b
huggingface.co
inclusionAI/SingGuard-0.8b-GGUF · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
📰 HuggingFace - Profiling in PyTorch (Part 3): Attention is all you profile
https://huggingface.co/blog/torch-attention-profile
https://huggingface.co/blog/torch-attention-profile
huggingface.co
Profiling in PyTorch (Part 3): Attention is all you profile
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
📰 OpenAI - How Deutsche Telekom is rewiring telecommunications with AI
https://openai.com/index/deutsche-telekom
https://openai.com/index/deutsche-telekom
OpenAI
How Deutsche Telekom is rewiring telecommunications with AI
How Deutsche Telekom is becoming an AI-native telco with OpenAI-transforming customer service, employee workflows, network operations, and the future of voice.
📰 PyTorch - Towards Free Normalization: Fusing Normalization into GEMM and Attention Kernels
Code available at: https://github.com/facebookresearch/ads_model_kernel_library/tree/main/multi_cta_norm_fusion and https://github.com/facebookresearch/ads_model_kernel_library/tree/main/gdpa_megakernel TL;DR In this blog post, we present various novel kernel fusion techniques for common normalization ops like LayerNorm and RMSNorm, which provide significant speedup...
https://pytorch.org/blog/towards-free-normalization-fusing-normalization-into-gemm-and-attention-kernels/
Code available at: https://github.com/facebookresearch/ads_model_kernel_library/tree/main/multi_cta_norm_fusion and https://github.com/facebookresearch/ads_model_kernel_library/tree/main/gdpa_megakernel TL;DR In this blog post, we present various novel kernel fusion techniques for common normalization ops like LayerNorm and RMSNorm, which provide significant speedup...
https://pytorch.org/blog/towards-free-normalization-fusing-normalization-into-gemm-and-attention-kernels/
GitHub
ads_model_kernel_library/multi_cta_norm_fusion at main · facebookresearch/ads_model_kernel_library
High-performance GPU kernels for Ads and Recsys model training, independently implemented and optimized for real-world workloads and model-specific input characteristics. - facebookresearch/ads_mod...
❤1
📰 Google AI Blog - LiteRT.js, Google's high performance Web AI Inference
We're excited to introduce LiteRT.js, the newest member of the LiteRT family! LiteRT.js is our powerful solution for running machine learning models directly in the browser, extending Google's cross-platform edge AI runtime to the web. Built for JavaScript developers, LiteRT.js delivers state-of-the-art ML model inference performance on WebGPU and upcoming WebNN, with a fallback to WebAssembly for CPU. This post provides a quick tour of LiteRT.js and gives web developers everything they need to get started.
https://developers.googleblog.com/en/litertjs-googles-high-performance-web-ai-inference/
We're excited to introduce LiteRT.js, the newest member of the LiteRT family! LiteRT.js is our powerful solution for running machine learning models directly in the browser, extending Google's cross-platform edge AI runtime to the web. Built for JavaScript developers, LiteRT.js delivers state-of-the-art ML model inference performance on WebGPU and upcoming WebNN, with a fallback to WebAssembly for CPU. This post provides a quick tour of LiteRT.js and gives web developers everything they need to get started.
https://developers.googleblog.com/en/litertjs-googles-high-performance-web-ai-inference/
Googleblog
Google for Developers Blog - News about Web, Mobile, AI and Cloud
Meet LiteRT.js: Google’s edge AI runtime for the web. Run ML models directly in the browser with high-performance WebGPU, WebNN, and WebAssembly.
📰 LMSys - Bringing DeepSeek-V4 Flash RL Training to AMD Instinct MI355X GPUs with Miles
https://lmsys.org/blog/2026-07-10-rocm-miles-dsv4
https://lmsys.org/blog/2026-07-10-rocm-miles-dsv4
www.lmsys.org
Bringing DeepSeek-V4 Flash RL Training to AMD Instinct MI355X GPUs with Miles
DeepSeek-V4 RL is now supported in Miles on AMD Instinct™ MI355X GPUs with ROCm™! RL requires SGLang rollout and Megatron training to implement the same policy closely enough that token probabilities ...
📰 Claude Blog - Working at the frontier: How Cognition trusts Claude Fable 5 to work through the night
https://claude.com/blog/working-at-the-frontier-how-cognition-trusts-claude-fable-5-to-work-through-the-night
https://claude.com/blog/working-at-the-frontier-how-cognition-trusts-claude-fable-5-to-work-through-the-night
Claude
Working at the frontier: How Cognition trusts Claude Fable 5 to work through the night | Claude by Anthropic
Cognition tested Claude Fable 5 in Devin, its AI software engineer. It's the first model its team trusts to run unattended for eight hours and deliver production-ready code.