π [HF Models] google - gemma-4-12B-it-qat-w4a16-ct
https://huggingface.co/google/gemma-4-12B-it-qat-w4a16-ct
π [HF Models] google - gemma-4-12B-it-qat-q4_0-gguf
https://huggingface.co/google/gemma-4-12B-it-qat-q4_0-gguf
π [HF Models] google - gemma-4-12B-it-qat-q4_0-unquantized-assistant
https://huggingface.co/google/gemma-4-12B-it-qat-q4_0-unquantized-assistant
π [HF Models] google - gemma-4-31B-it-qat-w4a16-ct
https://huggingface.co/google/gemma-4-31B-it-qat-w4a16-ct
π [HF Models] google - gemma-4-E2B-it-qat-q4_0-unquantized-assistant
https://huggingface.co/google/gemma-4-E2B-it-qat-q4_0-unquantized-assistant
π [HF Models] google - gemma-4-E4B-it-qat-q4_0-unquantized-assistant
https://huggingface.co/google/gemma-4-E4B-it-qat-q4_0-unquantized-assistant
π [HF Models] google - gemma-4-26B-A4B-it-qat-q4_0-unquantized-assistant
https://huggingface.co/google/gemma-4-26B-A4B-it-qat-q4_0-unquantized-assistant
π [HF Models] google - gemma-4-31B-it-qat-q4_0-unquantized-assistant
https://huggingface.co/google/gemma-4-31B-it-qat-q4_0-unquantized-assistant
π [HF Models] google - gemma-4-E2B-it-qat-mobile-ct
https://huggingface.co/google/gemma-4-E2B-it-qat-mobile-ct
π [HF Models] google - gemma-4-E4B-it-qat-mobile-ct
https://huggingface.co/google/gemma-4-E4B-it-qat-mobile-ct
π [HF Models] google - gemma-4-E2B-it-qat-mobile-transformers
https://huggingface.co/google/gemma-4-E2B-it-qat-mobile-transformers
π [HF Models] google - gemma-4-E4B-it-qat-mobile-transformers
https://huggingface.co/google/gemma-4-E4B-it-qat-mobile-transformers
π [HF Models] google - gemma-4-12B-it-qat-q4_0-unquantized
https://huggingface.co/google/gemma-4-12B-it-qat-q4_0-unquantized
π [HF Models] google - gemma-4-E4B-it-qat-w4a16-ct
https://huggingface.co/google/gemma-4-E4B-it-qat-w4a16-ct
https://huggingface.co/google/gemma-4-12B-it-qat-w4a16-ct
π [HF Models] google - gemma-4-12B-it-qat-q4_0-gguf
https://huggingface.co/google/gemma-4-12B-it-qat-q4_0-gguf
π [HF Models] google - gemma-4-12B-it-qat-q4_0-unquantized-assistant
https://huggingface.co/google/gemma-4-12B-it-qat-q4_0-unquantized-assistant
π [HF Models] google - gemma-4-31B-it-qat-w4a16-ct
https://huggingface.co/google/gemma-4-31B-it-qat-w4a16-ct
π [HF Models] google - gemma-4-E2B-it-qat-q4_0-unquantized-assistant
https://huggingface.co/google/gemma-4-E2B-it-qat-q4_0-unquantized-assistant
π [HF Models] google - gemma-4-E4B-it-qat-q4_0-unquantized-assistant
https://huggingface.co/google/gemma-4-E4B-it-qat-q4_0-unquantized-assistant
π [HF Models] google - gemma-4-26B-A4B-it-qat-q4_0-unquantized-assistant
https://huggingface.co/google/gemma-4-26B-A4B-it-qat-q4_0-unquantized-assistant
π [HF Models] google - gemma-4-31B-it-qat-q4_0-unquantized-assistant
https://huggingface.co/google/gemma-4-31B-it-qat-q4_0-unquantized-assistant
π [HF Models] google - gemma-4-E2B-it-qat-mobile-ct
https://huggingface.co/google/gemma-4-E2B-it-qat-mobile-ct
π [HF Models] google - gemma-4-E4B-it-qat-mobile-ct
https://huggingface.co/google/gemma-4-E4B-it-qat-mobile-ct
π [HF Models] google - gemma-4-E2B-it-qat-mobile-transformers
https://huggingface.co/google/gemma-4-E2B-it-qat-mobile-transformers
π [HF Models] google - gemma-4-E4B-it-qat-mobile-transformers
https://huggingface.co/google/gemma-4-E4B-it-qat-mobile-transformers
π [HF Models] google - gemma-4-12B-it-qat-q4_0-unquantized
https://huggingface.co/google/gemma-4-12B-it-qat-q4_0-unquantized
π [HF Models] google - gemma-4-E4B-it-qat-w4a16-ct
https://huggingface.co/google/gemma-4-E4B-it-qat-w4a16-ct
huggingface.co
google/gemma-4-12B-it-qat-w4a16-ct Β· Hugging Face
Weβre on a journey to advance and democratize artificial intelligence through open source and open science.
π° Google DeepMind - The latest AI news we announced in May 2026
https://blog.google/innovation-and-ai/technology/ai/google-ai-updates-may-2026/
https://blog.google/innovation-and-ai/technology/ai/google-ai-updates-may-2026/
Google
The latest AI news we announced in May 2026
Here are Googleβs latest AI updates from May 2026
π° Google AI Blog - Introducing the Google Colab CLI
Google has announced the Google Colab Command-Line Interface (CLI), a new tool that allows developers and AI agents to connect local terminals to remote Colab runtimes for frictionless execution. The lightweight CLI enables users to easily request high-powered GPUs, run local Python scripts remotely, and seamlessly retrieve artifact logs or models like fine-tuned Gemma 3 adapters. By integrating directly into standard terminal environments, the tool is highly programmable and ready to be used by AI agents such as Antigravity or Claude Code to manage complex machine learning pipelines.
https://developers.googleblog.com/en/introducing-the-google-colab-cli/
Google has announced the Google Colab Command-Line Interface (CLI), a new tool that allows developers and AI agents to connect local terminals to remote Colab runtimes for frictionless execution. The lightweight CLI enables users to easily request high-powered GPUs, run local Python scripts remotely, and seamlessly retrieve artifact logs or models like fine-tuned Gemma 3 adapters. By integrating directly into standard terminal environments, the tool is highly programmable and ready to be used by AI agents such as Antigravity or Claude Code to manage complex machine learning pipelines.
https://developers.googleblog.com/en/introducing-the-google-colab-cli/
Googleblog
Google for Developers Blog - News about Web, Mobile, AI and Cloud
Google announces the new Google Colab CLI, a lightweight tool bridging local terminals and remote runtimes for frictionless GPU/TPU offloading. Learn how developers and AI agents can execute remote scripts, download models, and automate ML pipelines.
β€1
π OpenAI - Biodefense in the Intelligence Age
https://openai.com/index/biodefense-in-the-intelligence-age
https://openai.com/index/biodefense-in-the-intelligence-age
OpenAI
Biodefense in the Intelligence Age
An action plan for AI-powered biological resilience
π° HuggingFace - Thousand Token Wood: shipping a multi-agent economy on a 3B model
https://huggingface.co/blog/build-small-hackathon/thousand-token-wood-sim
https://huggingface.co/blog/build-small-hackathon/thousand-token-wood-sim
huggingface.co
Thousand Token Wood: shipping a multi-agent economy on a 3B model
A Blog post by Build Small Hackathon on Hugging Face
π° HuggingFace - Persona Atlas: Mapping How Famous Minds Think
https://huggingface.co/blog/build-small-hackathon/persona-atlas
https://huggingface.co/blog/build-small-hackathon/persona-atlas
huggingface.co
Persona Atlas: Mapping How Famous Minds Think
A Blog post by Build Small Hackathon on Hugging Face
π [GitHub Releases] turboderp-org/exllamav3 - 0.0.40
https://github.com/turboderp-org/exllamav3/releases/tag/v0.0.40
https://github.com/turboderp-org/exllamav3/releases/tag/v0.0.40
GitHub
Release 0.0.40 Β· turboderp-org/exllamav3
An optimized quantization and inference library for running LLMs locally on modern consumer-class GPUs - Release 0.0.40 Β· turboderp-org/exllamav3
π° HuggingFace - Five labs, five minds: building a multi-model finance drama on small models
https://huggingface.co/blog/build-small-hackathon/thousand-token-wood-sim-v2
https://huggingface.co/blog/build-small-hackathon/thousand-token-wood-sim-v2
huggingface.co
Five labs, five minds: building a multi-model finance drama on small models
A Blog post by Build Small Hackathon on Hugging Face
π° Anthropic Research - Making Claude a chemist
https://www.anthropic.com/research/making-claude-a-chemist
https://www.anthropic.com/research/making-claude-a-chemist
Anthropic
Making Claude a chemist
Anthropic is working with world-class synthetic, computational, and analytical chemists to make Claude better at chemistry. In this post, we share our first work as part of this effort.
ποΈ Weekly GitHub Activity
π¦ llama.cpp
β Release: b9437 β b9544
β 107 commits
- Added support for EXAONE 4.5 #21733
- Added support for Granite4 Vision #23545
- Added support for Step3.7-Flash #23845
- Added support for Mellum architecture #23966
- Added support for Granite Multilingual Embeddings R2 #22716
- Added support for StepFun 3.5 MTP #23274
- Added tokenizer support for jina-embeddings-v2-base-zh #18756
- Initial support for Qwen3 SSM recurrent architectures #24031
- Server: Real-time reasoning interruption via new control endpoint #23971
- Server: Added placeholder bitmap for token counting and input_tokens API #23913
- Web UI: Thinking mode toggle, reasoning effort levels, and single-line preview #23434, #23601
- Web UI: Mermaid diagrams support and interactive preview #24032
- Tensor Parallel: Quantized KV cache support #23792
- Multimodal: Added frame merge support for Qwen-VL models #21858
- Vulkan: Optimized Q3_K/Q6_K performance on Intel Xe2/BMG via block loads 1962000
- CUDA: Improved MTP performance via mul_mat_vec_q_moe enrollment into PDL #24087
- Hexagon: Major optimizations for MUL_MAT, FLASH_ATTN, and GDN #23989
- Metal: Templated GLU kernels to support f16/f32 #23882
- Web UI: Custom CSS injection via configuration #23904
- KV-cache: SWA checkpoints store only non-masked cells #23981
- Deprecated llama_set_warmup #24009
- Fix model parameters not being propagated correctly to backend #23893
- Server: Avoid unnecessary checkpoint restore when new tokens are present #24110
- Fix session state corruption in common_prompt_batch_decode #23468
π All changes | Latest release
π¨ stable-diffusion.cpp
β Release: master-660-d2797b8 β master-679-f3fd359
β 19 commits
- Added support for Ideogram 4 models #1609
- Added support for Wan2.2 5B FLF2V #1110
- Implemented PiD support #1585
- Added Adaptive Projected Guidance (APG) and unconditional Skip Layer Guidance (SLG) #593
- Added --stream-layers to stream weights from CPU during generation #1576
- Added img-cfg support for edit models #929
- Optimized performance via pinned host buffer allocation and streaming budget management #1601 #1611
- Fixed Flash Attention KV padding issues #1453
π All changes | Latest release
π€ Fresh models trending on HuggingFace:
ideogram-ai/ideogram-4-nf4 β‘212
bosonai/higgs-audio-v3-tts-4b β‘153
Hcompany/Holo-3.1-4B β‘57
LiconStudio/LTX-2.3-Multiple-Subject-Reference β‘50
VAST-AI/TripoSplat β‘49
nex-agi/Nex-N2-Pro β‘48
Hcompany/Holo-3.1-35B-A3B β‘36
SupraLabs/Supra-50M-Reasoning β‘30
Aratako/Irodori-TTS-600M-v3-VoiceDesign β‘29
mudler/parakeet-cpp-gguf β‘28
nex-agi/Nex-N2-mini β‘22
Trendyol/Trendyol-TTS β‘22
litert-community/gemma-4-12B-it-litert-lm β‘20
latam-gpt/Llama-3.1-70B-LatamGPT-SFT-1.0 β‘20
Soul-AILab/SoulX-Transcriber β‘17
Hcompany/Holo-3.1-9B β‘17
Hcompany/Holo-3.1-0.8B β‘13
ideogram-ai/ideogram-4-nf4-diffusers β‘13
π¦ llama.cpp
β Release: b9437 β b9544
β 107 commits
- Added support for EXAONE 4.5 #21733
- Added support for Granite4 Vision #23545
- Added support for Step3.7-Flash #23845
- Added support for Mellum architecture #23966
- Added support for Granite Multilingual Embeddings R2 #22716
- Added support for StepFun 3.5 MTP #23274
- Added tokenizer support for jina-embeddings-v2-base-zh #18756
- Initial support for Qwen3 SSM recurrent architectures #24031
- Server: Real-time reasoning interruption via new control endpoint #23971
- Server: Added placeholder bitmap for token counting and input_tokens API #23913
- Web UI: Thinking mode toggle, reasoning effort levels, and single-line preview #23434, #23601
- Web UI: Mermaid diagrams support and interactive preview #24032
- Tensor Parallel: Quantized KV cache support #23792
- Multimodal: Added frame merge support for Qwen-VL models #21858
- Vulkan: Optimized Q3_K/Q6_K performance on Intel Xe2/BMG via block loads 1962000
- CUDA: Improved MTP performance via mul_mat_vec_q_moe enrollment into PDL #24087
- Hexagon: Major optimizations for MUL_MAT, FLASH_ATTN, and GDN #23989
- Metal: Templated GLU kernels to support f16/f32 #23882
- Web UI: Custom CSS injection via configuration #23904
- KV-cache: SWA checkpoints store only non-masked cells #23981
- Deprecated llama_set_warmup #24009
- Fix model parameters not being propagated correctly to backend #23893
- Server: Avoid unnecessary checkpoint restore when new tokens are present #24110
- Fix session state corruption in common_prompt_batch_decode #23468
π All changes | Latest release
π¨ stable-diffusion.cpp
β Release: master-660-d2797b8 β master-679-f3fd359
β 19 commits
- Added support for Ideogram 4 models #1609
- Added support for Wan2.2 5B FLF2V #1110
- Implemented PiD support #1585
- Added Adaptive Projected Guidance (APG) and unconditional Skip Layer Guidance (SLG) #593
- Added --stream-layers to stream weights from CPU during generation #1576
- Added img-cfg support for edit models #929
- Optimized performance via pinned host buffer allocation and streaming budget management #1601 #1611
- Fixed Flash Attention KV padding issues #1453
π All changes | Latest release
π€ Fresh models trending on HuggingFace:
ideogram-ai/ideogram-4-nf4 β‘212
bosonai/higgs-audio-v3-tts-4b β‘153
Hcompany/Holo-3.1-4B β‘57
LiconStudio/LTX-2.3-Multiple-Subject-Reference β‘50
VAST-AI/TripoSplat β‘49
nex-agi/Nex-N2-Pro β‘48
Hcompany/Holo-3.1-35B-A3B β‘36
SupraLabs/Supra-50M-Reasoning β‘30
Aratako/Irodori-TTS-600M-v3-VoiceDesign β‘29
mudler/parakeet-cpp-gguf β‘28
nex-agi/Nex-N2-mini β‘22
Trendyol/Trendyol-TTS β‘22
litert-community/gemma-4-12B-it-litert-lm β‘20
latam-gpt/Llama-3.1-70B-LatamGPT-SFT-1.0 β‘20
Soul-AILab/SoulX-Transcriber β‘17
Hcompany/Holo-3.1-9B β‘17
Hcompany/Holo-3.1-0.8B β‘13
ideogram-ai/ideogram-4-nf4-diffusers β‘13
GitHub
Add EXAONE 4.5 implementations by nuxlear Β· Pull Request #21733 Β· ggml-org/llama.cpp
Overview
Add support for the EXAONE 4.5 architecture for the EXAONE 4.5 model released by LG AI Research.
Additional information
This PR adds the modeling code for EXAONE 4.5, which uses the same...
Add support for the EXAONE 4.5 architecture for the EXAONE 4.5 model released by LG AI Research.
Additional information
This PR adds the modeling code for EXAONE 4.5, which uses the same...
π° HuggingFace - Her Β· ΰ€Ήΰ₯ΰ€° β a detective for your Claude Code sessions
https://huggingface.co/blog/build-small-hackathon/her-blog
https://huggingface.co/blog/build-small-hackathon/her-blog
π° HuggingFace - Sponsors especially OPENAI CODEX voucher usage for codex - openAI challange
https://huggingface.co/blog/build-small-hackathon/sponsors-vouchers
https://huggingface.co/blog/build-small-hackathon/sponsors-vouchers
huggingface.co
Sponsors especially OPENAI CODEX voucher usage for codex - openAI challange
A Blog post by Build Small Hackathon on Hugging Face
π° HuggingFace - Mythograph Atelier #1 - Abstract Art That Means Something to You
https://huggingface.co/blog/build-small-hackathon/mythograph-atelier-01-inspirations
https://huggingface.co/blog/build-small-hackathon/mythograph-atelier-01-inspirations
huggingface.co
Mythograph Atelier #1 - Abstract Art That Means Something to You
A Blog post by Build Small Hackathon on Hugging Face
π° Claude Blog - The Claude Cowork product guide
https://claude.com/blog/the-claude-cowork-product-guide
π° Claude Blog - How one Anthropic seller rebuilt his team's workflows with Claude Code
https://claude.com/blog/how-anthropic-uses-claude-gtm-engineering
https://claude.com/blog/the-claude-cowork-product-guide
π° Claude Blog - How one Anthropic seller rebuilt his team's workflows with Claude Code
https://claude.com/blog/how-anthropic-uses-claude-gtm-engineering
Claude
The Claude Cowork product guide | Claude by Anthropic
The Anthropic team shares how to get started with AnthriClaude Cowork, from setting up the tool to kicking off your first task.
π° HuggingFace - Amazing Digital Dentures (a failed project)
https://huggingface.co/blog/build-small-hackathon/amazingdigitaldentures
https://huggingface.co/blog/build-small-hackathon/amazingdigitaldentures
huggingface.co
Amazing Digital Dentures (a failed project)
A Blog post by Build Small Hackathon on Hugging Face
π° LMSys - Announcing the Recipient of the 2026 LMSYS PhD Fellowship
https://lmsys.org/blog/2026-06-08-lmsys-phd-fellowship
https://lmsys.org/blog/2026-06-08-lmsys-phd-fellowship
www.lmsys.org
Announcing the Recipient of the 2026 LMSYS PhD Fellowship - LMSYS Blog
We are delighted to announce the first recipient of the LMSYS Fellowship Program: Will Lin.
Following the launch of our Fellowship Program and careful review of applications, we selected Will for his...
Following the launch of our Fellowship Program and careful review of applications, we selected Will for his...
π° HuggingFace - Building Pakistan Notice Helper: A Small AI Tool for a Very Local Safety Problem
https://huggingface.co/blog/build-small-hackathon/building-pakistan-notice-helper
https://huggingface.co/blog/build-small-hackathon/building-pakistan-notice-helper
π° HuggingFace - The crash that vanished: control and emergence in a five-model economy
https://huggingface.co/blog/build-small-hackathon/thousand-token-wood-sim-v3
π° HuggingFace - The Open Source Community is backing OpenEnv for Agentic RL
https://huggingface.co/blog/openenv-agentic-rl
https://huggingface.co/blog/build-small-hackathon/thousand-token-wood-sim-v3
π° HuggingFace - The Open Source Community is backing OpenEnv for Agentic RL
https://huggingface.co/blog/openenv-agentic-rl
huggingface.co
The crash that vanished: control and emergence in a five-model economy
A Blog post by Build Small Hackathon on Hugging Face