📰 Anthropic - A research agenda for the Economic Futures Research Fund
https://www.anthropic.com/news/economic-futures-research-fund-agenda
📰 Anthropic - Ask Claude about the Anthropic Economic Index
https://www.anthropic.com/news/anthropic-economic-index-connector
📰 Anthropic - Anthropic is donating another $20 million to Public First Action
https://www.anthropic.com/news/donation-public-first-action
📰 Anthropic - Apply for Anthropic’s AI for Science rare disease research grants
https://www.anthropic.com/news/rare-disease-research-grants
https://www.anthropic.com/news/economic-futures-research-fund-agenda
📰 Anthropic - Ask Claude about the Anthropic Economic Index
https://www.anthropic.com/news/anthropic-economic-index-connector
📰 Anthropic - Anthropic is donating another $20 million to Public First Action
https://www.anthropic.com/news/donation-public-first-action
📰 Anthropic - Apply for Anthropic’s AI for Science rare disease research grants
https://www.anthropic.com/news/rare-disease-research-grants
Anthropic
A research agenda for the Economic Futures Research Fund
We’re committing $200 million to the Anthropic Economic Futures Research Fund to support ambitious external research.
📰 OpenAI - Building AI infrastructure with the Effingham County community
https://openai.com/index/building-ai-infrastructure-with-the-effingham-county-community
📰 OpenAI - How news organizations are using AI to advance their vital missions
https://openai.com/index/how-news-organizations-are-using-ai
📰 OpenAI - Advancing the next era of national science
https://openai.com/index/advancing-the-next-era-of-national-science
📰 OpenAI - Introducing OpenAI Presence
https://openai.com/index/introducing-openai-presence
📰 OpenAI - Introducing the ChatGPT for small business program
https://openai.com/index/introducing-chatgpt-small-business-program
📰 OpenAI - OpenAI and Hugging Face partner to address security incident during model evaluation
https://openai.com/index/hugging-face-model-evaluation-security-incident
📰 OpenAI - David Vélez and Robin Vince join the boards of the OpenAI Foundation and OpenAI Group PBC
https://openai.com/index/david-velez-robin-vince-join-openai-boards
https://openai.com/index/building-ai-infrastructure-with-the-effingham-county-community
📰 OpenAI - How news organizations are using AI to advance their vital missions
https://openai.com/index/how-news-organizations-are-using-ai
📰 OpenAI - Advancing the next era of national science
https://openai.com/index/advancing-the-next-era-of-national-science
📰 OpenAI - Introducing OpenAI Presence
https://openai.com/index/introducing-openai-presence
📰 OpenAI - Introducing the ChatGPT for small business program
https://openai.com/index/introducing-chatgpt-small-business-program
📰 OpenAI - OpenAI and Hugging Face partner to address security incident during model evaluation
https://openai.com/index/hugging-face-model-evaluation-security-incident
📰 OpenAI - David Vélez and Robin Vince join the boards of the OpenAI Foundation and OpenAI Group PBC
https://openai.com/index/david-velez-robin-vince-join-openai-boards
OpenAI
Building AI infrastructure with the Effingham County community
OpenAI announces Project Camellia in Effingham County, Georgia, with commitments to responsible energy, community investment, jobs, and access to Codex.
📰 NVIDIA - Make Long-Running NVIDIA TensorRT Engine Builds Observable and Cancelable in Python or C++
A TensorRT engine build can take seconds to many minutes. Large strongly typed models, deep tactic search, and a cold timing cache on a brand-new GPU SKU can…
https://developer.nvidia.com/blog/make-long-running-nvidia-tensorrt-engine-builds-observable-and-cancelable-in-python-or-c/
A TensorRT engine build can take seconds to many minutes. Large strongly typed models, deep tactic search, and a cold timing cache on a brand-new GPU SKU can…
https://developer.nvidia.com/blog/make-long-running-nvidia-tensorrt-engine-builds-observable-and-cancelable-in-python-or-c/
NVIDIA Technical Blog
Make Long-Running NVIDIA TensorRT Engine Builds Observable and Cancelable in Python or C++
A TensorRT engine build can take seconds to many minutes. Large strongly typed models, deep tactic search, and a cold timing cache on a brand-new GPU SKU can leave developers, end users…
📰 HuggingFace - Bringing Nunchaku 4-bit Diffusion Inference to Diffusers
https://huggingface.co/blog/nunchaku-diffusers
https://huggingface.co/blog/nunchaku-diffusers
huggingface.co
Bringing Nunchaku 4-bit Diffusion Inference to Diffusers
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
📰 PyTorch - Helion on TPU: Towards Hardware Heterogeneous Kernel Authoring
TL;DR Helion is PyTorch’s high-level DSL for writing performance-portable ML kernels. Partnering with Google, we have built a TPU backend that compiles Helion kernels to Pallas, providing a PyTorch-friendly way...
https://pytorch.org/blog/helion-on-tpu-towards-hardware-heterogeneous-kernel-authoring/
TL;DR Helion is PyTorch’s high-level DSL for writing performance-portable ML kernels. Partnering with Google, we have built a TPU backend that compiles Helion kernels to Pallas, providing a PyTorch-friendly way...
https://pytorch.org/blog/helion-on-tpu-towards-hardware-heterogeneous-kernel-authoring/
📰 OpenAI - Launching Health in ChatGPT
https://openai.com/index/health-in-chatgpt
🔓 OpenAI - NTT DATA Group cuts incident analysis to 30 minutes with Codex
https://openai.com/index/ntt-data
https://openai.com/index/health-in-chatgpt
🔓 OpenAI - NTT DATA Group cuts incident analysis to 30 minutes with Codex
https://openai.com/index/ntt-data
OpenAI
Launching Health in ChatGPT
Health in ChatGPT now lets eligible U.S. users securely connect medical records and Apple Health to get more personalized insights and better understand their health.
📰 NVIDIA - Start Customizing NVIDIA Nemotron 3 Nano with Prime Intellect Lab in Minutes
Customization is what enables developers to take a general model and tailor it to use cases, domains, languages, and more. However, customization comes with a…
https://developer.nvidia.com/blog/start-customizing-nvidia-nemotron-3-nano-with-prime-intellect-lab-in-minutes/
Customization is what enables developers to take a general model and tailor it to use cases, domains, languages, and more. However, customization comes with a…
https://developer.nvidia.com/blog/start-customizing-nvidia-nemotron-3-nano-with-prime-intellect-lab-in-minutes/
NVIDIA Technical Blog
Start Customizing NVIDIA Nemotron 3 Nano with Prime Intellect Lab in Minutes
Customization is what enables developers to take a general model and tailor it to use cases, domains, languages, and more. However, customization comes with a few challenges.
🆕 [HF Models] swiss-ai - Apertus-v1.5-70B
https://huggingface.co/swiss-ai/Apertus-v1.5-70B
🆕 [HF Models] swiss-ai - Apertus-v1.5-8B
https://huggingface.co/swiss-ai/Apertus-v1.5-8B
https://huggingface.co/swiss-ai/Apertus-v1.5-70B
🆕 [HF Models] swiss-ai - Apertus-v1.5-8B
https://huggingface.co/swiss-ai/Apertus-v1.5-8B
huggingface.co
swiss-ai/Apertus-v1.5-70B · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
📰 Google AI Blog - Run Ray on TPU, Part 2: Ray AI libraries
This second installment explores how Ray’s higher-level libraries—Serve, Data, and Train—abstract the complexities of running AI workloads on Google's TPU slices. Ray Serve uses a simple topology configuration to correctly gang-schedule large multi-host models, while Ray Data eliminates data-loading bottlenecks by feeding accelerators directly with native JAX batches. Finally, JaxTrainer streamlines distributed training across TPUs by automatically handling cross-slice coordination, checkpointing, and fault tolerance.
https://developers.googleblog.com/en/run-ray-on-tpu-part-2-ray-ai-libraries/
This second installment explores how Ray’s higher-level libraries—Serve, Data, and Train—abstract the complexities of running AI workloads on Google's TPU slices. Ray Serve uses a simple topology configuration to correctly gang-schedule large multi-host models, while Ray Data eliminates data-loading bottlenecks by feeding accelerators directly with native JAX batches. Finally, JaxTrainer streamlines distributed training across TPUs by automatically handling cross-slice coordination, checkpointing, and fault tolerance.
https://developers.googleblog.com/en/run-ray-on-tpu-part-2-ray-ai-libraries/
Googleblog
Google for Developers Blog - News about Web, Mobile, AI and Cloud
Learn how to scale AI workloads on TPU slices using Ray Serve for LLM deployment, Ray Data for fast JAX pipelines, and JaxTrainer for distributed training.
🔓 xAI - Bringing Grok 4.5 to iOS, Android, Web, and X
https://x.ai//news/grok-4-5-everywhere
🔓 xAI - Workflows in Grok Build
https://x.ai//news/workflows
🔓 xAI - Grok in Google Workspace
https://x.ai//news/introducing-google-workspace-addon
https://x.ai//news/grok-4-5-everywhere
🔓 xAI - Workflows in Grok Build
https://x.ai//news/workflows
🔓 xAI - Grok in Google Workspace
https://x.ai//news/introducing-google-workspace-addon
x.ai
Bringing Grok 4.5 to iOS, Android, Web, and X
Grok 4.5, our most intelligent model yet, is now on grok.com, X, iOS, and Android.
📰 NVIDIA - ModelExpress: Distributing Model Artifacts at the Speed of Light
Every byte moved has a cost. As model checkpoints grow to hundreds of gigabytes or even a terabyte, that cost adds up quickly. To make things even worse…
https://developer.nvidia.com/blog/modelexpress-distributing-model-artifacts-at-the-speed-of-light/
Every byte moved has a cost. As model checkpoints grow to hundreds of gigabytes or even a terabyte, that cost adds up quickly. To make things even worse…
https://developer.nvidia.com/blog/modelexpress-distributing-model-artifacts-at-the-speed-of-light/
NVIDIA Technical Blog
ModelExpress: Distributing Model Artifacts at the Speed of Light
Every byte moved has a cost. As model checkpoints grow to hundreds of gigabytes or even a terabyte, that cost adds up quickly. To make things even worse, moving these model weights around the cluster…
🔄 [GitHub Releases] sgl-project/sglang - v0.5.16
https://github.com/sgl-project/sglang/releases/tag/v0.5.16
https://github.com/sgl-project/sglang/releases/tag/v0.5.16
GitHub
Release v0.5.16 · sgl-project/sglang
Highlights
574 PRs from 169 contributors.
DSpark: confidence-driven speculative decoding: A new speculative algorithm. It drafts semi-autoregressively in blocks, then sizes each verify window from ...
574 PRs from 169 contributors.
DSpark: confidence-driven speculative decoding: A new speculative algorithm. It drafts semi-autoregressively in blocks, then sizes each verify window from ...
🔄 [GitHub Releases] vllm-project/vllm - v0.26.0
https://github.com/vllm-project/vllm/releases/tag/v0.26.0
https://github.com/vllm-project/vllm/releases/tag/v0.26.0
GitHub
Release v0.26.0 · vllm-project/vllm
vLLM v0.26.0 Release Notes
Highlights
This release features 411 commits from 212 contributors (61 new)!
New Inkling model family with a full support stack: base modeling (#48799), piecewise CUDA g...
Highlights
This release features 411 commits from 212 contributors (61 new)!
New Inkling model family with a full support stack: base modeling (#48799), piecewise CUDA g...