GenAI monitor
551 subscribers
4.33K links
AI frontier model updates & open source LLM releases
Download Telegram
๐Ÿ“ฐ Claude Blog - Working at the frontier: How Thomson Reuters builds AI for high-stakes professional work

https://claude.com/blog/working-at-the-frontier-how-thomson-reuters-builds-ai-for-high--stakes-professional-work


๐Ÿ“ฐ Claude Blog - How Anthropic's marketing operations team uses Claude Cowork to automate reporting and campaign builds

https://claude.com/blog/how-anthropics-marketing-operations-team-uses-claude-cowork-to-automate-reporting-and-campaign-builds


๐Ÿ“ฐ Claude Blog - Bringing Claude Code and Claude Cowork to government

https://claude.com/blog/bringing-claude-code-and-claude-cowork-to-government


๐Ÿ“ฐ Claude Blog - Choosing a Claude model and effort level in Claude Code

https://claude.com/blog/claude-model-and-effort-level-in-claude-code


๐Ÿ“ฐ Claude Blog - Claude Cowork is coming to mobile and web

https://claude.com/blog/cowork-web-mobile


๐Ÿ“ฐ Claude Blog - How people are using Claude Cowork

https://claude.com/blog/how-people-are-using-claude-cowork


๐Ÿ“ฐ Claude Blog - A field guide to Claude Fable 5: Finding your unknowns

https://claude.com/blog/a-field-guide-to-claude-fable-finding-your-unknowns
๐Ÿ“ฐ Anthropic - Inviting hard questions

https://www.anthropic.com/news/hard-questions


๐Ÿ“ฐ Anthropic - Ben Bernanke appointed to Anthropicโ€™s Long-Term Benefit Trust

https://www.anthropic.com/news/ben-bernanke


๐Ÿ“ฐ Anthropic - Introducing a way to reflect on how you use Claude

https://www.anthropic.com/news/reflect-with-claude
๐Ÿ†• [HF Models] nvidia - Kimi-K2.7-Code-DFlash

https://huggingface.co/nvidia/Kimi-K2.7-Code-DFlash


๐Ÿ”“ [HF Models] nvidia - ARDY-Core-RP-20FPS-Horizon40

https://huggingface.co/nvidia/ARDY-Core-RP-20FPS-Horizon40


๐Ÿ”“ [HF Models] nvidia - ARDY-G1-RP-25FPS-Horizon52

https://huggingface.co/nvidia/ARDY-G1-RP-25FPS-Horizon52


๐Ÿ”“ [HF Models] nvidia - ARDY-Core-RP-20FPS-Horizon8

https://huggingface.co/nvidia/ARDY-Core-RP-20FPS-Horizon8


๐Ÿ”“ [HF Models] nvidia - ARDY-G1-RP-25FPS-Horizon8

https://huggingface.co/nvidia/ARDY-G1-RP-25FPS-Horizon8
๐Ÿ†• [HF Models] inclusionAI - SingGuard-0.8b-GGUF

https://huggingface.co/inclusionAI/SingGuard-0.8b-GGUF


๐Ÿ”“ [HF Models] inclusionAI - SingGuard-2b

https://huggingface.co/inclusionAI/SingGuard-2b


๐Ÿ”“ [HF Models] inclusionAI - SingGuard-8b

https://huggingface.co/inclusionAI/SingGuard-8b


๐Ÿ”“ [HF Models] inclusionAI - SingGuard-4b

https://huggingface.co/inclusionAI/SingGuard-4b


๐Ÿ†• [HF Models] inclusionAI - SingGuard-8b-GGUF

https://huggingface.co/inclusionAI/SingGuard-8b-GGUF


๐Ÿ†• [HF Models] inclusionAI - SingGuard-4b-GGUF

https://huggingface.co/inclusionAI/SingGuard-4b-GGUF


๐Ÿ†• [HF Models] inclusionAI - SingGuard-2b-GGUF

https://huggingface.co/inclusionAI/SingGuard-2b-GGUF


๐Ÿ†• [HF Models] inclusionAI - SingGuard-0.8b

https://huggingface.co/inclusionAI/SingGuard-0.8b
๐Ÿ†• [HF Models] tencent - HiLS-Attention-7B


https://huggingface.co/tencent/HiLS-Attention-7B
๐Ÿ“ฐ Google AI Blog - LiteRT.js, Google's high performance Web AI Inference
We're excited to introduce LiteRT.js, the newest member of the LiteRT family! LiteRT.js is our powerful solution for running machine learning models directly in the browser, extending Google's cross-platform edge AI runtime to the web. Built for JavaScript developers, LiteRT.js delivers state-of-the-art ML model inference performance on WebGPU and upcoming WebNN, with a fallback to WebAssembly for CPU. This post provides a quick tour of LiteRT.js and gives web developers everything they need to get started.

https://developers.googleblog.com/en/litertjs-googles-high-performance-web-ai-inference/
๐Ÿ“ฐ NVIDIA - Reducing High-Bandwidth Memory Bottlenecks in JAX-Based LLM Training with Host Offloading
Large language model (LLM) training workloads increasingly run into GPU memory limits before compute is fully used. Model weights, gradients, optimizer statesโ€ฆ

https://developer.nvidia.com/blog/reducing-high-bandwidth-memory-bottlenecks-in-jax-based-llm-training-with-host-offloading/


๐Ÿ“ฐ NVIDIA - AI Model Co-Design: Hardware-Friendly LLM Design
AI performance comes down to three dimensions: Deployments must balance all three: High accuracy is wasted if responses are slow, and raw throughput meansโ€ฆ

https://developer.nvidia.com/blog/ai-model-co-design-hardware-friendly-llm-design/


๐Ÿ“ฐ NVIDIA - Accelerating End-to-End Co-Folding Performance with NVIDIA BioNeMo Agent Toolkit
Biomolecular structure prediction and co-folding with models like OpenFold3 are now mainstream, large-scale workloads powering drug discovery and protein design.

https://developer.nvidia.com/blog/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit/
๐Ÿ†• [HF Models] openbmb - UltraX-0.6B-Preview


https://huggingface.co/openbmb/UltraX-0.6B-Preview