📰 HuggingFace - Five labs, five minds: building a multi-model finance drama on small models
https://huggingface.co/blog/build-small-hackathon/thousand-token-wood-sim-v2
https://huggingface.co/blog/build-small-hackathon/thousand-token-wood-sim-v2
huggingface.co
Five labs, five minds: building a multi-model finance drama on small models
A Blog post by Build Small Hackathon on Hugging Face
📰 Anthropic Research - Making Claude a chemist
https://www.anthropic.com/research/making-claude-a-chemist
https://www.anthropic.com/research/making-claude-a-chemist
Anthropic
Making Claude a chemist
Anthropic is working with world-class synthetic, computational, and analytical chemists to make Claude better at chemistry. In this post, we share our first work as part of this effort.
🗓️ Weekly GitHub Activity
🦙 llama.cpp
└ Release: b9437 → b9544
└ 107 commits
- Added support for EXAONE 4.5 #21733
- Added support for Granite4 Vision #23545
- Added support for Step3.7-Flash #23845
- Added support for Mellum architecture #23966
- Added support for Granite Multilingual Embeddings R2 #22716
- Added support for StepFun 3.5 MTP #23274
- Added tokenizer support for jina-embeddings-v2-base-zh #18756
- Initial support for Qwen3 SSM recurrent architectures #24031
- Server: Real-time reasoning interruption via new control endpoint #23971
- Server: Added placeholder bitmap for token counting and input_tokens API #23913
- Web UI: Thinking mode toggle, reasoning effort levels, and single-line preview #23434, #23601
- Web UI: Mermaid diagrams support and interactive preview #24032
- Tensor Parallel: Quantized KV cache support #23792
- Multimodal: Added frame merge support for Qwen-VL models #21858
- Vulkan: Optimized Q3_K/Q6_K performance on Intel Xe2/BMG via block loads 1962000
- CUDA: Improved MTP performance via mul_mat_vec_q_moe enrollment into PDL #24087
- Hexagon: Major optimizations for MUL_MAT, FLASH_ATTN, and GDN #23989
- Metal: Templated GLU kernels to support f16/f32 #23882
- Web UI: Custom CSS injection via configuration #23904
- KV-cache: SWA checkpoints store only non-masked cells #23981
- Deprecated llama_set_warmup #24009
- Fix model parameters not being propagated correctly to backend #23893
- Server: Avoid unnecessary checkpoint restore when new tokens are present #24110
- Fix session state corruption in common_prompt_batch_decode #23468
🔗 All changes | Latest release
🎨 stable-diffusion.cpp
└ Release: master-660-d2797b8 → master-679-f3fd359
└ 19 commits
- Added support for Ideogram 4 models #1609
- Added support for Wan2.2 5B FLF2V #1110
- Implemented PiD support #1585
- Added Adaptive Projected Guidance (APG) and unconditional Skip Layer Guidance (SLG) #593
- Added --stream-layers to stream weights from CPU during generation #1576
- Added img-cfg support for edit models #929
- Optimized performance via pinned host buffer allocation and streaming budget management #1601 #1611
- Fixed Flash Attention KV padding issues #1453
🔗 All changes | Latest release
🤗 Fresh models trending on HuggingFace:
ideogram-ai/ideogram-4-nf4 ♡212
bosonai/higgs-audio-v3-tts-4b ♡153
Hcompany/Holo-3.1-4B ♡57
LiconStudio/LTX-2.3-Multiple-Subject-Reference ♡50
VAST-AI/TripoSplat ♡49
nex-agi/Nex-N2-Pro ♡48
Hcompany/Holo-3.1-35B-A3B ♡36
SupraLabs/Supra-50M-Reasoning ♡30
Aratako/Irodori-TTS-600M-v3-VoiceDesign ♡29
mudler/parakeet-cpp-gguf ♡28
nex-agi/Nex-N2-mini ♡22
Trendyol/Trendyol-TTS ♡22
litert-community/gemma-4-12B-it-litert-lm ♡20
latam-gpt/Llama-3.1-70B-LatamGPT-SFT-1.0 ♡20
Soul-AILab/SoulX-Transcriber ♡17
Hcompany/Holo-3.1-9B ♡17
Hcompany/Holo-3.1-0.8B ♡13
ideogram-ai/ideogram-4-nf4-diffusers ♡13
🦙 llama.cpp
└ Release: b9437 → b9544
└ 107 commits
- Added support for EXAONE 4.5 #21733
- Added support for Granite4 Vision #23545
- Added support for Step3.7-Flash #23845
- Added support for Mellum architecture #23966
- Added support for Granite Multilingual Embeddings R2 #22716
- Added support for StepFun 3.5 MTP #23274
- Added tokenizer support for jina-embeddings-v2-base-zh #18756
- Initial support for Qwen3 SSM recurrent architectures #24031
- Server: Real-time reasoning interruption via new control endpoint #23971
- Server: Added placeholder bitmap for token counting and input_tokens API #23913
- Web UI: Thinking mode toggle, reasoning effort levels, and single-line preview #23434, #23601
- Web UI: Mermaid diagrams support and interactive preview #24032
- Tensor Parallel: Quantized KV cache support #23792
- Multimodal: Added frame merge support for Qwen-VL models #21858
- Vulkan: Optimized Q3_K/Q6_K performance on Intel Xe2/BMG via block loads 1962000
- CUDA: Improved MTP performance via mul_mat_vec_q_moe enrollment into PDL #24087
- Hexagon: Major optimizations for MUL_MAT, FLASH_ATTN, and GDN #23989
- Metal: Templated GLU kernels to support f16/f32 #23882
- Web UI: Custom CSS injection via configuration #23904
- KV-cache: SWA checkpoints store only non-masked cells #23981
- Deprecated llama_set_warmup #24009
- Fix model parameters not being propagated correctly to backend #23893
- Server: Avoid unnecessary checkpoint restore when new tokens are present #24110
- Fix session state corruption in common_prompt_batch_decode #23468
🔗 All changes | Latest release
🎨 stable-diffusion.cpp
└ Release: master-660-d2797b8 → master-679-f3fd359
└ 19 commits
- Added support for Ideogram 4 models #1609
- Added support for Wan2.2 5B FLF2V #1110
- Implemented PiD support #1585
- Added Adaptive Projected Guidance (APG) and unconditional Skip Layer Guidance (SLG) #593
- Added --stream-layers to stream weights from CPU during generation #1576
- Added img-cfg support for edit models #929
- Optimized performance via pinned host buffer allocation and streaming budget management #1601 #1611
- Fixed Flash Attention KV padding issues #1453
🔗 All changes | Latest release
🤗 Fresh models trending on HuggingFace:
ideogram-ai/ideogram-4-nf4 ♡212
bosonai/higgs-audio-v3-tts-4b ♡153
Hcompany/Holo-3.1-4B ♡57
LiconStudio/LTX-2.3-Multiple-Subject-Reference ♡50
VAST-AI/TripoSplat ♡49
nex-agi/Nex-N2-Pro ♡48
Hcompany/Holo-3.1-35B-A3B ♡36
SupraLabs/Supra-50M-Reasoning ♡30
Aratako/Irodori-TTS-600M-v3-VoiceDesign ♡29
mudler/parakeet-cpp-gguf ♡28
nex-agi/Nex-N2-mini ♡22
Trendyol/Trendyol-TTS ♡22
litert-community/gemma-4-12B-it-litert-lm ♡20
latam-gpt/Llama-3.1-70B-LatamGPT-SFT-1.0 ♡20
Soul-AILab/SoulX-Transcriber ♡17
Hcompany/Holo-3.1-9B ♡17
Hcompany/Holo-3.1-0.8B ♡13
ideogram-ai/ideogram-4-nf4-diffusers ♡13
GitHub
Add EXAONE 4.5 implementations by nuxlear · Pull Request #21733 · ggml-org/llama.cpp
Overview
Add support for the EXAONE 4.5 architecture for the EXAONE 4.5 model released by LG AI Research.
Additional information
This PR adds the modeling code for EXAONE 4.5, which uses the same...
Add support for the EXAONE 4.5 architecture for the EXAONE 4.5 model released by LG AI Research.
Additional information
This PR adds the modeling code for EXAONE 4.5, which uses the same...
📰 HuggingFace - Her · हेर — a detective for your Claude Code sessions
https://huggingface.co/blog/build-small-hackathon/her-blog
https://huggingface.co/blog/build-small-hackathon/her-blog
📰 HuggingFace - Sponsors especially OPENAI CODEX voucher usage for codex - openAI challange
https://huggingface.co/blog/build-small-hackathon/sponsors-vouchers
https://huggingface.co/blog/build-small-hackathon/sponsors-vouchers
huggingface.co
Sponsors especially OPENAI CODEX voucher usage for codex - openAI challange
A Blog post by Build Small Hackathon on Hugging Face
📰 HuggingFace - Mythograph Atelier #1 - Abstract Art That Means Something to You
https://huggingface.co/blog/build-small-hackathon/mythograph-atelier-01-inspirations
https://huggingface.co/blog/build-small-hackathon/mythograph-atelier-01-inspirations
huggingface.co
Mythograph Atelier #1 - Abstract Art That Means Something to You
A Blog post by Build Small Hackathon on Hugging Face
📰 Claude Blog - The Claude Cowork product guide
https://claude.com/blog/the-claude-cowork-product-guide
📰 Claude Blog - How one Anthropic seller rebuilt his team's workflows with Claude Code
https://claude.com/blog/how-anthropic-uses-claude-gtm-engineering
https://claude.com/blog/the-claude-cowork-product-guide
📰 Claude Blog - How one Anthropic seller rebuilt his team's workflows with Claude Code
https://claude.com/blog/how-anthropic-uses-claude-gtm-engineering
Claude
The Claude Cowork product guide | Claude by Anthropic
The Anthropic team shares how to get started with AnthriClaude Cowork, from setting up the tool to kicking off your first task.
📰 HuggingFace - Amazing Digital Dentures (a failed project)
https://huggingface.co/blog/build-small-hackathon/amazingdigitaldentures
https://huggingface.co/blog/build-small-hackathon/amazingdigitaldentures
huggingface.co
Amazing Digital Dentures (a failed project)
A Blog post by Build Small Hackathon on Hugging Face
📰 LMSys - Announcing the Recipient of the 2026 LMSYS PhD Fellowship
https://lmsys.org/blog/2026-06-08-lmsys-phd-fellowship
https://lmsys.org/blog/2026-06-08-lmsys-phd-fellowship
www.lmsys.org
Announcing the Recipient of the 2026 LMSYS PhD Fellowship - LMSYS Blog
We are delighted to announce the first recipient of the LMSYS Fellowship Program: Will Lin.
Following the launch of our Fellowship Program and careful review of applications, we selected Will for his...
Following the launch of our Fellowship Program and careful review of applications, we selected Will for his...
📰 HuggingFace - Building Pakistan Notice Helper: A Small AI Tool for a Very Local Safety Problem
https://huggingface.co/blog/build-small-hackathon/building-pakistan-notice-helper
https://huggingface.co/blog/build-small-hackathon/building-pakistan-notice-helper
📰 HuggingFace - The crash that vanished: control and emergence in a five-model economy
https://huggingface.co/blog/build-small-hackathon/thousand-token-wood-sim-v3
📰 HuggingFace - The Open Source Community is backing OpenEnv for Agentic RL
https://huggingface.co/blog/openenv-agentic-rl
https://huggingface.co/blog/build-small-hackathon/thousand-token-wood-sim-v3
📰 HuggingFace - The Open Source Community is backing OpenEnv for Agentic RL
https://huggingface.co/blog/openenv-agentic-rl
huggingface.co
The crash that vanished: control and emergence in a five-model economy
A Blog post by Build Small Hackathon on Hugging Face
📰 OpenAI - Introducing the OpenAI Economic Research Exchange
https://openai.com/index/economic-research-exchange
📰 OpenAI - OpenAI Economic Research Exchange
https://openai.com/form/economic-research-exchange
https://openai.com/index/economic-research-exchange
📰 OpenAI - OpenAI Economic Research Exchange
https://openai.com/form/economic-research-exchange
OpenAI
Introducing the OpenAI Economic Research Exchange
OpenAI launches the Economic Research Exchange to study AI’s impact on jobs, productivity, and the economy. Applications are now open for selected research projects.
📰 NVIDIA - Train Models Faster with JAX and MaxText Using NVFP4 on NVIDIA Blackwell
Pre-training frontier LLMs comes down to throughput. When training spans trillions of tokens across thousands of accelerators, every percentage point of step…
https://developer.nvidia.com/blog/train-models-faster-with-jax-and-maxtext-using-nvfp4-on-nvidia-blackwell/
Pre-training frontier LLMs comes down to throughput. When training spans trillions of tokens across thousands of accelerators, every percentage point of step…
https://developer.nvidia.com/blog/train-models-faster-with-jax-and-maxtext-using-nvfp4-on-nvidia-blackwell/
NVIDIA Technical Blog
Train Models Faster with JAX and MaxText Using NVFP4 on NVIDIA Blackwell
Pre-training frontier LLMs comes down to throughput. When training spans trillions of tokens across thousands of accelerators, every percentage point of step time can add up to days of training and…
🔄 [GitHub Releases] Dao-AILab/flash-attention - v2.8.3.post1
https://github.com/Dao-AILab/flash-attention/releases/tag/v2.8.3.post1
https://github.com/Dao-AILab/flash-attention/releases/tag/v2.8.3.post1
GitHub
Release v2.8.3.post1 · Dao-AILab/flash-attention
Fast and memory-efficient exact attention. Contribute to Dao-AILab/flash-attention development by creating an account on GitHub.
📰 HuggingFace - NeuroBait: I fine-tuned a model to spark dopamine for ADHD brain
https://huggingface.co/blog/build-small-hackathon/neurobait-adhd
https://huggingface.co/blog/build-small-hackathon/neurobait-adhd
huggingface.co
NeuroBait: I fine-tuned a model to spark dopamine for ADHD brain
A Blog post by Build Small Hackathon on Hugging Face