🚨 AI News | TestingCatalog
7.58K subscribers
4.33K photos
704 videos
40 files
4.33K links
Latest AI News on AI Agents, Model Releases, Tools, Leaks, and Rumors πŸ—ž
Download Telegram
ANTHROPIC πŸ”₯: Claude Code now can pull project instructions from AGENTS md file in case if CLAUDE md file is not present.

A code cleanup on a global scale πŸ‘€

It is inevitable that most of the common coding AI agents will be eventually working on same repositories and will have to find ways on how to better collaborate together. Standardisation is always good.
🀣107❀5😴1
SPACEXAI πŸ”₯: Grok Voice Transcribe 2.0, a new multilingual speech-to-text model, has been announced.

> Grok Voice Transcribe 2.0 ranks first for accuracy among 32 streaming models on the Artificial Analysis leaderboard

> Grok Voice Transcribe 2.0 supports batch and streaming, speaker diarization, multichannel transcription, key term biasing, text formatting, filler word removal, and smart turn detection.
❀6πŸ‘2πŸ”₯2
Meta announced Muse Connectors, enabling developers and businesses to submit their own via a Muse platform.

Notion, Granola and other Connectors are now available.

Meta has partnered with Stripe to make payment possible.

β€œReach new customers and give existing ones more ways to use your product.”


I feel that Muse has a chance to get into a very nice spot by being an AI assistant for consumers and small businesses because it does exactly what they all want.

AI agent payments will be a new norm, and Meta grub a huge chunk of this market. Especially in a combination with Meta Glasses.

Side note, Muse is now available in Canada!
❀6πŸ‘€2πŸ‘1πŸ”₯1
SpaceXAI says Grok Voice Transcribe 2 doubles accuracy

SpaceXAI launched Grok Voice Transcribe 2.0, claiming 2x the accuracy of v1.0 at unchanged pricing. It supports multilingual, real-time and batch transcription, with diarization, timestamps, term biasing, and phone-audio gains.

πŸ—ž #spacexai @testingcatalog
πŸ‘4❀1πŸ”₯1
OPENAI πŸ”₯: A new banked reset for Codex and ChatGPT Work users is coming. A big ship on Tuesday is confirmed too!

I feel like all labs should adopt this rule - no ships equal β€œreset”.

True? πŸ‘€
πŸ’―12😱8❀3πŸ‘Ž2
StepFun released Step 5 Preview, a new 600B total / 27B active MoE model with 1M context.

> Step 5 Preview comes with vision capabilities and proficiency in software engineering and finance tasks.

> It is now available in StepFun AI Studio and APIs.

Pareto is a new front line πŸ‘€
❀6πŸ‘2😐2πŸ‘€2
Anthropic tests Fable 5.2 and Opus 5.5 ahead of the release

Anthropic appears to be quietly A/B testing Claude Fable 5.2, with users reporting stronger one-shot JavaScript outputs and richer canvas demos. Early comparisons suggest a clear step up from 5.1, as Opus 5.5 rumors also build.

πŸ—ž #anthropic @testingcatalog
πŸ‘146πŸ”₯2
Anthropic has rebranded Workbench as Playground! It has been in development for quite a while and is now more conventional.

> "Workbench is now Playground - Explore the Claude API by using it. Test models and Messages API features without writing code."

> "Start from working examples for tools, citations, and caching
See each request and response, with tokens, cost, and latency
Copy the SDK code when something works"
4❀3
SpaceXAI πŸ”₯: Grok 4.7 has been released on Grok Build, Cursor and APIs, offering close to Opus 5 performance on certain tasks at a price of Grok 4.6.

Testing time πŸ‘€
❀3πŸ‘€1