OPENAI π₯: A close look into an upcoming Work Louder's Codex Micro.
> Monitor the status of your Codex agents in a satisfyingly new tactile way.
> Your new AI agent controller.
> 6 RGB agent status keys, customizable command keys via the Input software, and 6 programmable layers with app-based auto-switching. h/t x@btibor91
This controller needs a Reset button too π
> Monitor the status of your Codex agents in a satisfyingly new tactile way.
> Your new AI agent controller.
> 6 RGB agent status keys, customizable command keys via the Input software, and 6 programmable layers with app-based auto-switching. h/t x@btibor91
This controller needs a Reset button too π
β€11π3π₯1 1
OpenAI prepares Codex Micro keypad to control AI Agents
OpenAI launched Codex Micro, a Work Louder keypad built for Codex. Its 6 RGB Agent Keys show thread status in real time, while remappable controls and app layers would support coding workflows beyond Codex.
π #leak @testingcatalog
OpenAI launched Codex Micro, a Work Louder keypad built for Codex. Its 6 RGB Agent Keys show thread status in real time, while remappable controls and app layers would support coding workflows beyond Codex.
π #leak @testingcatalog
TestingCatalog AI News
OpenAI prepares Codex Micro keypad to control AI Agents
OpenAI launches the Codex Micro keypad, bringing real-time RGB agent status and customizable controls to developers managing multiple AI agents.
β€4π1
Media is too big
VIEW IN TELEGRAM
Kivine, a new model on Arena, is potentially an upcoming Kimi K3 model.
I got lucky and caught a comparison between Fable 5 and Kimi K3 in the Universe simulation prompt.
> Fable 5 finished faster, and most UX components were more robust and easy to use.
> Kimi K3 was much more complex and visually appealing. At some point, if you select a planet and set the speed to x100, it will spin you around in the 1st-person pov.
It seems like they are very close π
I got lucky and caught a comparison between Fable 5 and Kimi K3 in the Universe simulation prompt.
> Fable 5 finished faster, and most UX components were more robust and easy to use.
> Kimi K3 was much more complex and visually appealing. At some point, if you select a planet and set the speed to x100, it will spin you around in the 1st-person pov.
It seems like they are very close π
β€7 3
π¨ AI News | TestingCatalog
OpenAI prepares Codex Micro keypad to control AI Agents OpenAI launched Codex Micro, a Work Louder keypad built for Codex. Its 6 RGB Agent Keys show thread status in real time, while remappable controls and app layers would support coding workflows beyondβ¦
Media is too big
VIEW IN TELEGRAM
OPENAI π₯: Codex Micro is now officially available on the OpenAI Supply website for $230!
> Designed with Work Louder, the kbd-1.0-codex-micro brings your agent workspace into reach. Keep active chats close, spot what every agent is doing through live RGB feedback, and map your most-used Codex actions to tactile controls built for the way you actually ship.
The device is available in two options: Silent and Clicky!
> Designed with Work Louder, the kbd-1.0-codex-micro brings your agent workspace into reach. Keep active chats close, spot what every agent is doing through live RGB feedback, and map your most-used Codex actions to tactile controls built for the way you actually ship.
The device is available in two options: Silent and Clicky!
β€4π2π2
Thinking Machines released Inkling, an open-weight model, available on a new Inkling Playground in the Tinker console.
> Mixture-of-Experts transformer with 975B total parameters and 41B active parameters.
> It supports a context window of up to 1M tokens.
> Inkling scored 1257 points on Design Arenaβs Agentic Web Dev leaderboard (GPT-5.6-Sol got 1260).
> Mixture-of-Experts transformer with 975B total parameters and 41B active parameters.
> It supports a context window of up to 1M tokens.
> Inkling scored 1257 points on Design Arenaβs Agentic Web Dev leaderboard (GPT-5.6-Sol got 1260).
β€6π2
OpenAI announced GPT-Red, an internal model for finding prompt-injection vulnerabilities at scale.
> GPTβRed is a strong red-teamer, and our previous models are highly vulnerable to its prompt injection attacks.
> We use GPTβRed to adversarially train GPTβ5.6, making it much more robust to prompt injections.
> GPTβRed is a strong red-teamer, and our previous models are highly vulnerable to its prompt injection attacks.
> We use GPTβRed to adversarially train GPTβ5.6, making it much more robust to prompt injections.
β€5π2 1
Thinking Machines debuts open-weight Inkling AI model
Thinking Machines Lab released Inkling, an open-weights multimodal MoE model with 975B parameters, 1M-token context, and tunable reasoning. It supports customization, broad deployment options, and strong coding, reasoning, vision, and audio results.
π #ai @testingcatalog
Thinking Machines Lab released Inkling, an open-weights multimodal MoE model with 975B parameters, 1M-token context, and tunable reasoning. It supports customization, broad deployment options, and strong coding, reasoning, vision, and audio results.
π #ai @testingcatalog
TestingCatalog AI News
Thinking Machines debuts open-weight Inkling AI model
Thinking Machines Lab launches Inkling, its open-weight Mixture-of-Experts AI model, offering broad customization and multi-modal support.
π₯4β€2π2
This media is not supported in your browser
VIEW IN TELEGRAM
Kimi K3 has been teased officially π
β€8π5π₯2
Anthropic reset 5h and weekly rate limits on Claude for all users.
3 more Fable 5 days left (until July 19) and it likely wonβt come back.
Last testing chance π
3 more Fable 5 days left (until July 19) and it likely wonβt come back.
Last testing chance π
β€10π8π3 2
Early look at Kimi K3 generations from Moonshot AI on Arena
Moonshot AI appears close to launching Kimi K3, with leaks, teasers, and test sightings pointing to an imminent debut. Early reports suggest frontier-level performance, strong 3D and coding output, but slower runtimes on complex tasks.
π #kimi @testingcatalog
Moonshot AI appears close to launching Kimi K3, with leaks, teasers, and test sightings pointing to an imminent debut. Early reports suggest frontier-level performance, strong 3D and coding output, but slower runtimes on complex tasks.
π #kimi @testingcatalog
TestingCatalog AI News
Early look at Kimi K3 generations from Moonshot AI on Arena
Moonshot AIβs Kimi K3 launch seems imminent, with leaks and teasers hinting at a groundbreaking multi-trillion-parameter MoE model.
β€6π3
MOONSHOT π₯: Kimi K3 from Moonshot AI landed on the first place on Frontend Code Arena with 1679 points, outperforming Claude Fable 5 and GPT-5.6 Sol by far.
β€12π₯3π2
Meta β€οΈ OpenRouter
Muse Spark 1.1 is now available on OpenRouter for US users.
Muse Spark 1.1 is now available on OpenRouter for US users.
β€5π3