Early look at Kimi K3 generations from Moonshot AI on Arena
Moonshot AI appears close to launching Kimi K3, with leaks, teasers, and test sightings pointing to an imminent debut. Early reports suggest frontier-level performance, strong 3D and coding output, but slower runtimes on complex tasks.
π #kimi @testingcatalog
Moonshot AI appears close to launching Kimi K3, with leaks, teasers, and test sightings pointing to an imminent debut. Early reports suggest frontier-level performance, strong 3D and coding output, but slower runtimes on complex tasks.
π #kimi @testingcatalog
TestingCatalog AI News
Early look at Kimi K3 generations from Moonshot AI on Arena
Moonshot AIβs Kimi K3 launch seems imminent, with leaks and teasers hinting at a groundbreaking multi-trillion-parameter MoE model.
β€6π3
MOONSHOT π₯: Kimi K3 from Moonshot AI landed on the first place on Frontend Code Arena with 1679 points, outperforming Claude Fable 5 and GPT-5.6 Sol by far.
β€12π₯3π2
Meta β€οΈ OpenRouter
Muse Spark 1.1 is now available on OpenRouter for US users.
Muse Spark 1.1 is now available on OpenRouter for US users.
β€5π3
This media is not supported in your browser
VIEW IN TELEGRAM
Google is working on a native menu for Skills on Gemini desktop. These Skills will potentially become available in all chats (and not only in Spark as it is now).
Users will be able to upload, create and edit their Gemini Skills, as well as use Gemini to create them.
Skill folders will be available there too, allowing users to select local folders containing necessary Skills.
Users will be able to upload, create and edit their Gemini Skills, as well as use Gemini to create them.
Skill folders will be available there too, allowing users to select local folders containing necessary Skills.
β€8π4
ICYMI π: Grok Heavy subscribers now get X Premium+ for free!
Would be cool for X Premium Business subscription to include Grok Heavy one day.
Would be cool for X Premium Business subscription to include Grok Heavy one day.
β€5π4
OPENAI π₯: ChatGPT desktop app now includes ChatGPT chats and projects too! The history is being synced across desktop, web, and mobile.
The Chat and Work mode switcher now works consistently with the web version, and Codex mode continues to function as before.
A very nice change! π
The Chat and Work mode switcher now works consistently with the web version, and Codex mode continues to function as before.
A very nice change! π
π₯10π4β€3
X now has a dedicated AI timeline to follow latest AI model releases and more!
To enable it, go to Add+ > Add Timeline > Select βArtificial Intelligenceβ > Profit!
Even though my βFor Youβ timeline is exactly the same, βAIβ is now my default.
Cya there π
To enable it, go to Add+ > Add Timeline > Select βArtificial Intelligenceβ > Profit!
Even though my βFor Youβ timeline is exactly the same, βAIβ is now my default.
Cya there π
1β€11 4π2
Perplexity launches SPACE runtime for AI agent tasks
Perplexityβs SPACE is a VM-based runtime for long-running AI agents, separating sessions from disposable sandboxes. It supports pause, resume, rollback, and recovery, while cutting sandbox startup times by 3.1 to 5x in production.
π #perplexity @testingcatalog
Perplexityβs SPACE is a VM-based runtime for long-running AI agents, separating sessions from disposable sandboxes. It supports pause, resume, rollback, and recovery, while cutting sandbox startup times by 3.1 to 5x in production.
π #perplexity @testingcatalog
TestingCatalog AI News
Perplexity launches SPACE runtime for AI agent tasks
Perplexity unveils SPACE, a new sandboxed runtime that powers all Computer sessions and enables persistent AI agent operations at scale.
1β€4π2
OpenAI unveils GPT-Red to boost AI safety with internal red-teaming
OpenAI introduced GPT-Red, an internal red-team model that finds prompt-injection flaws at scale. Trained via self-play, it outperformed humans in attacks and has already reduced failure rates in newer GPT models through adversarial training.
π #openai @testingcatalog
OpenAI introduced GPT-Red, an internal red-team model that finds prompt-injection flaws at scale. Trained via self-play, it outperformed humans in attacks and has already reduced failure rates in newer GPT models through adversarial training.
π #openai @testingcatalog
TestingCatalog AI News
OpenAI unveils GPT-Red to boost AI safety with internal red-teaming
OpenAI introduces GPTβRed, a specialized internal system for uncovering vulnerabilities in AI models using advanced red-teaming and self-play learning.
π4β€3π₯1
ANTHROPIC π₯: Claude Fable 5 will only be included in Max and Team Premium plans, starting from July 20 at 50% of limits.
> Pro and Team Standard users will receive a one-time $100 credit and have access to Fable 5 via credits only.
> We're continuing to invest in new capacity and will keep everyone updated as we do.
Compute is all you need π€
> Pro and Team Standard users will receive a one-time $100 credit and have access to Fable 5 via credits only.
> We're continuing to invest in new capacity and will keep everyone updated as we do.
Compute is all you need π€
π8π₯7β€5
SPACEXAI π₯: The next 2T params Grok model is expected to finish training next week and supposed to exceed Kimi K3 in performance with a better speed and token efficiency.
If we will be able to see this model in September, that would mean that SpaceXAI managed to put this process on the right rails.
If we will be able to see this model in September, that would mean that SpaceXAI managed to put this process on the right rails.
β€12π6 3