ByteDance launches SeedRealtime full-duplex AI model
ByteDanceβs SeedRealtime is a full-duplex audio-visual LLM that processes speech, video, and text in one model, reducing latency and context loss while handling timing, scene cues, and real-world dialogue with fewer pacing errors.
π #bytedance @testingcatalog
ByteDanceβs SeedRealtime is a full-duplex audio-visual LLM that processes speech, video, and text in one model, reducing latency and context loss while handling timing, scene cues, and real-world dialogue with fewer pacing errors.
π #bytedance @testingcatalog
TestingCatalog AI News
ByteDance launches SeedRealtime full-duplex AI model
ByteDance has fully rolled out SeedRealtime, a full-duplex model that watches, listens and speaks across continuous audio-visual streams.
π4β€1
BREAKING π₯: Meta AI released Muse Code, a new terminal coding agent powered by the newest Muse Spark 1.2!
It scores 59% on DeepSWE 1.1, outperforming Grok Build 4.5 and Gemini 3.6 Flash.
It scores 59% on DeepSWE 1.1, outperforming Grok Build 4.5 and Gemini 3.6 Flash.
β€7π3π2
Meta launches Muse Code beta powered by Muse Spark 1.2
Meta launched Muse Code beta, a terminal agent for full software tasks across large repositories on macOS and Linux. Powered by Muse Spark 1.2, it plans, codes, validates, logs work locally, and supports restart-safe sessions.
π #meta @testingcatalog
Meta launched Muse Code beta, a terminal agent for full software tasks across large repositories on macOS and Linux. Powered by Muse Spark 1.2, it plans, codes, validates, logs work locally, and supports restart-safe sessions.
π #meta @testingcatalog
TestingCatalog AI News
Meta launches Muse Code beta powered by Muse Spark 1.2
Muse Code is now in beta on macOS and Linux, using persistent background agents to plan, implement and validate complex changes across large repositories.
π2
OpenAI is planning to launch ads on ChatGPT in Brazil and Mexico soon.
Besides that, OpenAI released a plenty of improvements to its ads platform, including conversion optimisation feature and carousel ad formats.
ADGI expansion π€
Besides that, OpenAI released a plenty of improvements to its ads platform, including conversion optimisation feature and carousel ad formats.
ADGI expansion π€
β€6π3π1
Prime Intellect announced Prime Agent, a new self-improving RLM harness for coding and long-running autonomous tasks.
Prime Agent scored 95.5% on ARC-AGI-3! π€―
> The Recursive Language Model (RLM) treats context as a variable and subagent delegation as function calls inside a REPL.
> Continual Harness treats the harness's own state, abstracted as its prompts, skills, memory, and sub-agents, as something the agent can create, read, update, and delete (CRUD) from its own trajectory.
The next level of abstraction π
Prime Agent scored 95.5% on ARC-AGI-3! π€―
> The Recursive Language Model (RLM) treats context as a variable and subagent delegation as function calls inside a REPL.
> Continual Harness treats the harness's own state, abstracted as its prompts, skills, memory, and sub-agents, as something the agent can create, read, update, and delete (CRUD) from its own trajectory.
The next level of abstraction π
β€6π₯6π4π1
Grok voice mode now supports connectors!
This enables Grok to execute a wide range of tasks, bridging the gap with Claude and others.
Rolling out gradually π
This enables Grok to execute a wide range of tasks, bridging the gap with Claude and others.
Rolling out gradually π
β€6π3π3
Sesame Preview is now available on Android.
One of the best voice modes available on the market and my βdrive modeβ default.
One of the best voice modes available on the market and my βdrive modeβ default.
β€6π3π₯1
Loads of changes happening in Google and the hope for Gemini 3.5 Pro is slowly evaporating. However, the team is committed to focus on frontier intelligence.
Gemini 4 WIP π
> We are committed to being at the frontier and are super focused on the areas where we need to improve.
> The Gemini models are in good hands with Koray and the leads, and I'm excited about the great progress weβre making with our new models, including Gemini 4.
Gemini 4 WIP π
> We are committed to being at the frontier and are super focused on the areas where we need to improve.
> The Gemini models are in good hands with Koray and the leads, and I'm excited about the great progress weβre making with our new models, including Gemini 4.
β€9π₯4π1π€ͺ1
DeepSeek is planning to βSignificantlyβ increase their API pricing soon as per new notice on DeepSeek platform.
No more cheap tokens π
No more cheap tokens π
π€¬12β€5π€―3π2
OPENAI π₯: ChatGPT Free users are getting access to GPT-5.6 Luna along with a new Think button for harder questions.
Plus and Pro subscribers are getting an updated GPT-5.6 Sol as well as a new thinking effort slider!
> Plus and Pro users can access the updated version of GPTβ5.6 Sol and the new slider in ChatGPT starting today.
> GPTβ5.6 Luna will become the default model for Free and Go users this week.
Plus and Pro subscribers are getting an updated GPT-5.6 Sol as well as a new thinking effort slider!
> Plus and Pro users can access the updated version of GPTβ5.6 Sol and the new slider in ChatGPT starting today.
> GPTβ5.6 Luna will become the default model for Free and Go users this week.
β€11π6
BREAKING π₯: The upcoming OpenAI device will be a doughnut shaped smart speaker with interactive components, according to Bloomberg.
> A new device from OpenAI will have a unique look, with moving parts and a doughnut shape, and will likely cost more than $300.
> The device, a smart speaker without a display, is designed to be easy to carry around and will have features such as speaker grills, microphones, and lights to demonstrate when it's listening.
> OpenAI plans to release the device in 2027, and it will be positioned as an AI-first computer that can help users get things done, with the goal of making it feel more alive than current smart speakers.
I need to test it π
> A new device from OpenAI will have a unique look, with moving parts and a doughnut shape, and will likely cost more than $300.
> The device, a smart speaker without a display, is designed to be easy to carry around and will have features such as speaker grills, microphones, and lights to demonstrate when it's listening.
> OpenAI plans to release the device in 2027, and it will be positioned as an AI-first computer that can help users get things done, with the goal of making it feel more alive than current smart speakers.
I need to test it π
β€8π3
ByteDance is pre-training a ~10T parameters model, according to Financial Times.
Given the fact how easily they took over on video generation, we may see some big surprises in a general use area as well.
The end of the year will be hot π₯
Given the fact how easily they took over on video generation, we may see some big surprises in a general use area as well.
The end of the year will be hot π₯
β€9π₯3