📰 NVIDIA - Train Models Faster with JAX and MaxText Using NVFP4 on NVIDIA Blackwell
Pre-training frontier LLMs comes down to throughput. When training spans trillions of tokens across thousands of accelerators, every percentage point of step…
https://developer.nvidia.com/blog/train-models-faster-with-jax-and-maxtext-using-nvfp4-on-nvidia-blackwell/
Pre-training frontier LLMs comes down to throughput. When training spans trillions of tokens across thousands of accelerators, every percentage point of step…
https://developer.nvidia.com/blog/train-models-faster-with-jax-and-maxtext-using-nvfp4-on-nvidia-blackwell/
NVIDIA Technical Blog
Train Models Faster with JAX and MaxText Using NVFP4 on NVIDIA Blackwell
Pre-training frontier LLMs comes down to throughput. When training spans trillions of tokens across thousands of accelerators, every percentage point of step time can add up to days of training and…
🔄 [GitHub Releases] Dao-AILab/flash-attention - v2.8.3.post1
https://github.com/Dao-AILab/flash-attention/releases/tag/v2.8.3.post1
https://github.com/Dao-AILab/flash-attention/releases/tag/v2.8.3.post1
GitHub
Release v2.8.3.post1 · Dao-AILab/flash-attention
Fast and memory-efficient exact attention. Contribute to Dao-AILab/flash-attention development by creating an account on GitHub.
📰 HuggingFace - NeuroBait: I fine-tuned a model to spark dopamine for ADHD brain
https://huggingface.co/blog/build-small-hackathon/neurobait-adhd
https://huggingface.co/blog/build-small-hackathon/neurobait-adhd
huggingface.co
NeuroBait: I fine-tuned a model to spark dopamine for ADHD brain
A Blog post by Build Small Hackathon on Hugging Face
🆕 [HF Models] jdopensource - JoyAI-Image-Edit-ComfyUI
https://huggingface.co/jdopensource/JoyAI-Image-Edit-ComfyUI
https://huggingface.co/jdopensource/JoyAI-Image-Edit-ComfyUI
📰 Google Model Cards - Gemini 3.5 Audio (Live Translate)
https://deepmind.google/models/model-cards/gemini-3-5-audio/
https://deepmind.google/models/model-cards/gemini-3-5-audio/
Google DeepMind
Gemini 3.5 Audio (Live Translate) - Model Card
📰 Anthropic Research - Paving the way for agents in biology
https://www.anthropic.com/research/agents-in-biology
https://www.anthropic.com/research/agents-in-biology
Anthropic
Paving the way for agents in biology
In this Anthropic Science post, Laura Luebbert argues that we need to make biological data infrastructure more agent-friendly.
📰 Anthropic - Claude Fable 5 and Claude Mythos 5
https://www.anthropic.com/news/claude-fable-5-mythos-5
https://www.anthropic.com/news/claude-fable-5-mythos-5
Anthropic
Claude Fable 5 and Claude Mythos 5
Today we’re launching Claude Fable 5: a Mythos-class model that we’ve made safe for general use.
📰 HuggingFace - Can Voice Agents Handle Bilingual Customers? Benchmarking Frontier ASR on Code-Switched Speech
https://huggingface.co/blog/ServiceNow-AI/code-switching
https://huggingface.co/blog/ServiceNow-AI/code-switching
📰 OpenAI - How engineers at Nextdoor use Codex to build without limits
https://openai.com/index/nextdoor
📰 OpenAI - Confidential submission of draft S-1 to the SEC
https://openai.com/index/openai-submits-confidential-s-1
📰 OpenAI - Built to benefit everyone: our plan
https://openai.com/index/built-to-benefit-everyone-our-plan
📰 OpenAI - OpenAI Champion Programs
https://openai.com/academy/champion-programs
https://openai.com/index/nextdoor
📰 OpenAI - Confidential submission of draft S-1 to the SEC
https://openai.com/index/openai-submits-confidential-s-1
📰 OpenAI - Built to benefit everyone: our plan
https://openai.com/index/built-to-benefit-everyone-our-plan
📰 OpenAI - OpenAI Champion Programs
https://openai.com/academy/champion-programs
OpenAI
How engineers at Nextdoor use Codex to build without limits
How engineers at Nextdoor use Codex with GPT-5.5 to investigate hard-to-reproduce issues, build across platforms, and focus on product outcomes.
🔓 HuggingFace - Migrating Your GitHub CI to Hugging Face Jobs
https://huggingface.co/blog/github-ci-hf-jobs
🔓 HuggingFace - Introducing North Mini Code: Cohere’s First Model For Developers
https://huggingface.co/blog/CohereLabs/introducing-north-mini-code
https://huggingface.co/blog/github-ci-hf-jobs
🔓 HuggingFace - Introducing North Mini Code: Cohere’s First Model For Developers
https://huggingface.co/blog/CohereLabs/introducing-north-mini-code
huggingface.co
Migrating Your GitHub CI to Hugging Face Jobs
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
📰 NVIDIA - Delivering Lifecycle Control for AI Infrastructure at Scale with NVIDIA DGX Spark Enterprise Manageability
As AI infrastructure scales, enterprise expectations for operational maturity are increasing. Organizations expect these systems to be provisionable, observable…
https://developer.nvidia.com/blog/delivering-lifecycle-control-for-ai-infrastructure-at-scale-with-nvidia-dgx-spark-enterprise-manageability/
📰 NVIDIA - Model Quantization: Turn FP8 Checkpoints into High-Performance Inference Engines with NVIDIA TensorRT
Converting a quantized checkpoint into an NVIDIA TensorRT engine bridges the gap between model optimization and production deployment, enabling faster inference…
https://developer.nvidia.com/blog/model-quantization-turn-fp8-checkpoints-into-high-performance-inference-engines-with-nvidia-tensorrt/
📰 NVIDIA - Accelerating Federated Learning Research with AI Agents and NVIDIA FLARE Auto-FL
Federated learning (FL) research often begins with a deceptively simple question: What should we try next? A new aggregation rule, a FedProx coefficient…
https://developer.nvidia.com/blog/accelerating-federated-learning-research-with-ai-agents-and-nvidia-flare-auto-fl/
📰 NVIDIA - Evaluate Clinical ASR Models Faster with Agent Skills and NVIDIA Nemotron Speech
Training a speech AI model to correctly recognize or synthesize clinical terminology is surprisingly difficult. Drug names like Acetaminophen, Amlodipine…
https://developer.nvidia.com/blog/evaluate-clinical-asr-models-faster-with-agent-skills-and-nvidia-nemotron-speech/
As AI infrastructure scales, enterprise expectations for operational maturity are increasing. Organizations expect these systems to be provisionable, observable…
https://developer.nvidia.com/blog/delivering-lifecycle-control-for-ai-infrastructure-at-scale-with-nvidia-dgx-spark-enterprise-manageability/
📰 NVIDIA - Model Quantization: Turn FP8 Checkpoints into High-Performance Inference Engines with NVIDIA TensorRT
Converting a quantized checkpoint into an NVIDIA TensorRT engine bridges the gap between model optimization and production deployment, enabling faster inference…
https://developer.nvidia.com/blog/model-quantization-turn-fp8-checkpoints-into-high-performance-inference-engines-with-nvidia-tensorrt/
📰 NVIDIA - Accelerating Federated Learning Research with AI Agents and NVIDIA FLARE Auto-FL
Federated learning (FL) research often begins with a deceptively simple question: What should we try next? A new aggregation rule, a FedProx coefficient…
https://developer.nvidia.com/blog/accelerating-federated-learning-research-with-ai-agents-and-nvidia-flare-auto-fl/
📰 NVIDIA - Evaluate Clinical ASR Models Faster with Agent Skills and NVIDIA Nemotron Speech
Training a speech AI model to correctly recognize or synthesize clinical terminology is surprisingly difficult. Drug names like Acetaminophen, Amlodipine…
https://developer.nvidia.com/blog/evaluate-clinical-asr-models-faster-with-agent-skills-and-nvidia-nemotron-speech/
NVIDIA Technical Blog
Delivering Lifecycle Control for AI Infrastructure at Scale with NVIDIA DGX Spark Enterprise Manageability
As AI infrastructure scales, enterprise expectations for operational maturity are increasing. Organizations expect these systems to be provisionable, observable, secure, and manageable at scale—the…
🆕 [HF Models] google - diffusiongemma-26B-A4B-it
https://huggingface.co/google/diffusiongemma-26B-A4B-it
https://huggingface.co/google/diffusiongemma-26B-A4B-it
📰 PyTorch - Portable vLLM Model Inference Kernels in Helion
TL;DR Helion kernels were integrated into vLLM for FP8 inference using Qwen3 models and evaluated across NVIDIA H100 and B200 GPUs. The experiments show that Helion provides a productive PyTorch-native...
https://pytorch.org/blog/portable-vllm-model-inference-kernels-in-helion/
TL;DR Helion kernels were integrated into vLLM for FP8 inference using Qwen3 models and evaluated across NVIDIA H100 and B200 GPUs. The experiments show that Helion provides a productive PyTorch-native...
https://pytorch.org/blog/portable-vllm-model-inference-kernels-in-helion/