🚨 AI News | TestingCatalog
7.53K subscribers
4.28K photos
688 videos
40 files
4.31K links
Latest AI News on AI Agents, Model Releases, Tools, Leaks, and Rumors πŸ—ž
Download Telegram
OPENAI πŸ”₯: New GPT-Live-Transcribe and GPT-Transcribe models are now available on APIs!

> GPT-Live-Transcribe - low-latency live transcription.

> GPT-Transcribe - asynchronous transcription of completed audio files and batch workloads.
❀4πŸ”₯2πŸ‘1
GOOGLE πŸ”₯: Gemini Notebook is about to get even smarter! In addition, users will be able to control whether watermarks are added to generated artifacts.

"Gemini Notebook just got smarter. Try asking to find new sources from the web"
πŸ‘8❀551
Cursor launches Start plan in India for β‚Ή649 per month

Cursor Start launches in India at β‚Ή649/month with tax included, adding Grok 4.5, Composer, more agent requests, and UPI support. The tier targets daily development across desktop, web, iOS, and CLI.

πŸ—ž #cursor @testingcatalog
❀5πŸ‘1
Google is working on interactive Apps for Gemini Notebook

Google is testing an App artifact for Gemini Notebook that could turn source material into runnable apps from prompts. The feature is not live, but aligns with an earlier secure code sandbox release.

πŸ—ž #notebooklm @testingcatalog
❀5πŸ”₯2
🚨 AI News | TestingCatalog
Google is working on interactive Apps for Gemini Notebook Google is testing an App artifact for Gemini Notebook that could turn source material into runnable apps from prompts. The feature is not live, but aligns with an earlier secure code sandbox release.…
Media is too big
VIEW IN TELEGRAM
GOOGLE πŸ”₯: A new "App" artifact is being developed on Gemini Notebook!

> "Generate an interactive HTML app based on your sources"

Users will be able to customize it via a prompt; examples include interactive dashboards, apps, and games! Imagine playing a game based on books you have added as sources?!

Besides that, Google continues to work on AI Notes and may also add an option to enable or disable watermarking for generated artifacts.
❀922
Bagel Labs launches WorldDiT world model for robotics

Bagel Labs released WorldDiT, an open robotics world model that jointly predicts robot actions and future scenes with one diffusion backbone, delivering strong LIBERO performance at sub-billion scale for on-robot deployment.

πŸ—ž #ai @testingcatalog
3❀2
SpaceXAI launches Grok Voice Think Fast 2.0 on Agent Builder

xAI launched Grok Voice Think Fast 2.0, a speech model for voice agents with higher benchmark scores, faster first audio, stronger transcription across 24 languages, and lower latency. It costs $0.08 per audio minute.

πŸ—ž #spacexai @testingcatalog
❀2πŸ”₯2πŸ‘1
GOOGLE πŸ”₯: Lyria 3.5 has been released on Google Flow Music! Besides that, Flow Music now has covers, lip-sync videos, and an iOS app.

> Meet Lyria 3.5. Experience dynamic vocals, richer musicality, and advanced creative controls with our new flagship model.

> Reimagine your music with Covers. Transform songs into a completely new style while keeping the original structure intact.

> Take the studio with you. Download the Google Flow Music iOS app to create, listen, and share from anywhere.

> Direct lip-synced music videos. Use Gemini Omni Flash and new lip-syncing capabilities to create stunning visuals.
❀9πŸ‘5
This media is not supported in your browser
VIEW IN TELEGRAM
CURSOR πŸ”₯: The latest version of Cursor for iOS is now compatible with iPad! Besides that, the app got a new Inbox and a complete PR review experience.

I need more devices for testing πŸ‘€
❀4πŸ‘4πŸ‘Ž1
Revolut 🀝 OpenAI

Revolut introduced ChatGPT Go subscription plan as a benefit for their customers.

> Starting from 3 months free on Standard, you can unlock up to 12 months of ChatGPT Go included at no extra cost, depending on your plan.

ChatGPT user base is about to grow quite a lot soon!
❀8πŸ‘3πŸ”₯3
GOOGLE πŸ”₯: Gemini desktop app for macOS now supports "Speak to Window" feature!

> "Today, we’re introducing a new way to create, edit, and summarize with Gemini using just your voice, right where you are already working."
πŸ‘4❀3
Google introduced Gemini Robotics ER 2, a new embodied reasoning model!

Benchmarks πŸ‘€

> Success/failure detection: Now operates on raw video feeds rather than static snapshots to catch mid-execution failures like spills, slips, or misalignments.

> General instrument reading: Extends beyond circular dials and sight glasses to include digital displays, linear scales, rulers, and liquid thermometers. We tested it across 10 different types of instruments.

> Enhanced spatial VQA: Improves Visual Question Answering throughGemini’s advancements in multi-modal understanding.
❀6πŸ”₯1πŸ€”1
OPENAI πŸ”₯: GPT-5.6 Luna prices got reduced by 80% along with a 20% cut for GPT-5.6 Terra!

GPT-5.6 Sol got a new faster option on the API with a 2.5x speed boost at 2x price.

Many models got overshadowed πŸ‘€
❀9πŸ‘3πŸ¦„2🀩1
THINKING MACHINES πŸ”₯: A new open-weight model has been released!

Inkling-Small has 276B total parameters, 12B active, and is now available on Huggingface and Tinker Playground.

And it is quite fast πŸ‘€
❀9πŸ₯°3πŸ‘2🀩2
Codex got a new "activity" sidebar in the latest update! The button works as a toggle to switch between "Projects & Recent chats" and "Activity" views.

> See chats that are unread, active or awaiting response.
❀8πŸ‘3πŸ”₯3πŸ‘1
Media is too big
VIEW IN TELEGRAM
PERPLEXITY πŸ”₯: Spaces got upgraded to Projects, a new type of workspaces for collaboration with Perplexity Computer, powered by a shared file system and self-improving Brain memory!

> Between tasks, Brain reviews the Project's files and sessions and updates what it knows, so each task starts with full context from previous work.
❀5πŸ‘5
Dreamina Seedance 2.5 is now available on Dreamina AI for paid plans in Southeast Asia, the Middle East, Africa, Europe, and South America.

No US πŸ‘€

- Native 30s videos
- A new interactive editing experience
- Long video mode (up to 3 minutes)
- Dreamina AI plugins for Maya and Blender
- Up to 50 multimodal references
- More true-to-life lighting and shadows
❀9πŸ”₯2😴1
DEEPSEEK πŸ”₯: A new DeepSeek-V4-Flash-0731 model has been released on APIs in Beta!

> DeepSeek-V4-Flash-0731 surpasses V4-Pro-Preview on agentic benchmarks.

> V4-Flash now supports the Responses API format.

Additionally, DeepSeek-V4-Pro is expected to be released ASAP!
❀15πŸ”₯2
MiniMax H3 is now available on HailuoAI & MiniMax APIs!

> All-in-One Reference - Creating Videos Using Text, Image, Audio, and Video Inputs

> Precise editing controls - Improve and iterate by adhering to precise guidelines.

> Built for all creative scenarios - from movies and commercials to games, brands, and e-commerce

> 2K video from $0.081/sec, 768p video coming soon, from $0.047/sec
❀6πŸ”₯3😴2