OpenAI tests next-gen Image V2 model on ChatGPT and LM Arena
OpenAI is testing Image V2 on LM Arena in three variants, showing stronger UI rendering, accurate text, and prompt fidelity. Some users access it via A/B tests. A launch could counter Googleβs lead, though pricing and quality retention remain uncertain.
π #chatgpt
OpenAI is testing Image V2 on LM Arena in three variants, showing stronger UI rendering, accurate text, and prompt fidelity. Some users access it via A/B tests. A launch could counter Googleβs lead, though pricing and quality retention remain uncertain.
π #chatgpt
TestingCatalog
OpenAI tests next-gen Image V2 model on ChatGPT and LM Arena
OpenAI is quietly testing its next-gen ImageV2 model on LM Arena, with early testers noting strong prompt accuracy and realistic UI rendering.
π3π₯1
Google prepares Jules V2 agent capable of taking bigger tasks
Google is reportedly developing βJitro,β a next-gen Jules coding agent centered on goal-driven development. Instead of task-based prompts, it would pursue KPI-level objectives within a dedicated workspace, signaling a shift toward autonomous, outcome-focused AI collaboration for engineering teams.
π #jules
Google is reportedly developing βJitro,β a next-gen Jules coding agent centered on goal-driven development. Instead of task-based prompts, it would pursue KPI-level objectives within a dedicated workspace, signaling a shift toward autonomous, outcome-focused AI collaboration for engineering teams.
π #jules
TestingCatalog
Google prepares Jules V2 agent capable of taking bigger tasks
Google is reportedly developing Jitro, the next-gen Jules coding agent, shifting to KPI-driven AI coding assistance with autonomous goal-setting.
π3π₯3β€1 1
Atomic Bot now runs local AI models on your computer
Atomic Bot now runs fully offline via Ollama, letting users deploy OpenClaw on local hardware with no API keys or cloud reliance. It supports major open models, 40,000+ skills, and multi-messenger automation, packaged as a native macOS and Windows app.
π #sponsored
Atomic Bot now runs fully offline via Ollama, letting users deploy OpenClaw on local hardware with no API keys or cloud reliance. It supports major open models, 40,000+ skills, and multi-messenger automation, packaged as a native macOS and Windows app.
π #sponsored
TestingCatalog
Atomic Bot now runs local AI models on your computer
Atomic Bot now runs OpenClaw on local models with no API keys or tokens. Your personal AI assistant, fully offline on your machine.
π4
Telegram adds AI text editor and upgraded polls for all users
Telegram has launched an update featuring a Cocoon AI-powered Text Editor that corrects, translates, and restyles messages with privacy protection and no data retention. The release also upgrades polls with media support, custom options, quiz modes, and advanced admin controls.
π #telegram
Telegram has launched an update featuring a Cocoon AI-powered Text Editor that corrects, translates, and restyles messages with privacy protection and no data retention. The release also upgrades polls with media support, custom options, quiz modes, and advanced admin controls.
π #telegram
TestingCatalog
Telegram adds AI text editor and upgraded polls for all users
What's new? Telegram's AI editor powered by Cocoon AI corrects grammar and translates texts; Telegram added polls with media, location and quiz options;
π8π₯2 2 1
Anthropic likely tests managed 24/7 AI agents for Businesses
Anthropicβs Conway project is evolving into an enterprise-focused agent platform with extensions, browser control, and cryptographically verified webhooks. Its design supports secure system-to-system automation, positioning Conway as managed infrastructure for teams building workflows around Claude.
π #claude
Anthropicβs Conway project is evolving into an enterprise-focused agent platform with extensions, browser control, and cryptographically verified webhooks. Its design supports secure system-to-system automation, positioning Conway as managed infrastructure for teams building workflows around Claude.
π #claude
TestingCatalog
Anthropic likely tests managed 24/7 AI agents for Businesses
Anthropicβs Conway project adds support for extensions, browser tasks, and secure webhooks, signaling a shift to business use cases.
β€2π₯2 1 1
Anthropic signed an agreement with Google and Broadcom for multiple gigawatts of next-generation TPU capacity.
With all whatβs coming from Anthropic it feels very much needed. Also, managed 24/7 agents will consume a lot.
With all whatβs coming from Anthropic it feels very much needed. Also, managed 24/7 agents will consume a lot.
π7π₯4β€3
Looks like Google started preparing Projects on Gemini for the upcoming rollout. At this moment, it is likely an unintended appearance.
The time has come π
h/t https://t.me/c/1349477688/33765
The time has come π
h/t https://t.me/c/1349477688/33765
β€6π5π₯3 1
π¨ AI News | TestingCatalog
BREAKING π¨: ANTHROPIC IS WORKING ON ITS OWN ALWAYS-ON AGENT SOLUTION CALLED CONWAY! CONWAY WILL HAVE A SEPARATE UI INSTANCE, WILL BE ABLE TO OPERATE BROWSER, CONNECTORS, CLAUDE CODE (EPITAXY?) AND COULD BE INVOKED VIA WEBHOOKS. IT WILL ALSO SUPPORT EXTENSIONSβ¦
BREAKING π¨: Anthropic is working on adding Claude Conway always-on agent to its mobile app.
This means that we potentially may see a consumer version too!
Soon? π
This means that we potentially may see a consumer version too!
Soon? π
β€6 1
Some users are noticing a new layout being rolled out on DeepSeek along with a potential stealth V4 model update.
Have you seen it too? π
Have you seen it too? π
π8β€4 2
BREAKING π¨: ANTHROPIC ANNOUNCED CYBERSECURITY PROJECT GLASSWING AND MYTHOS BENCHMARKS!
Claude Mythos scored 93.9% on SWE Bench Verified and 87.3 on SWE Bench Multilingual!
βWe do not plan to make Claude Mythos Preview generally available, but our eventual goal is to enable our users to safely deploy Mythos-class models at scaleβ
Claude Mythos scored 93.9% on SWE Bench Verified and 87.3 on SWE Bench Multilingual!
βWe do not plan to make Claude Mythos Preview generally available, but our eventual goal is to enable our users to safely deploy Mythos-class models at scaleβ
β€5π₯5π2
Anthropic announces Claude Mythos for cybersecurity research
Anthropic introduced Claude Mythos Preview, an AI model that autonomously detects and exploits zero-day vulnerabilities. It has uncovered critical flaws across major systems and is available to select partners, with $100 million in credits supporting cybersecurity efforts.
π #claude
Anthropic introduced Claude Mythos Preview, an AI model that autonomously detects and exploits zero-day vulnerabilities. It has uncovered critical flaws across major systems and is available to select partners, with $100 million in credits supporting cybersecurity efforts.
π #claude
TestingCatalog
Anthropic announces Claude Mythos for cybersecurity research
What's new? AnthropiC unveiled Claude Mythos Preview to spot zero-day flaws and craft exploits; select partners get access via Claude API for limited testing;
β€4π1
Zhipu AI launches open-source GLM-5.1 model for coding tasks
Z AI has launched GLM-5.1, a flagship model built for agentic engineering and long-horizon coding, capable of running up to eight hours on a single task.
π #ai
Z AI has launched GLM-5.1, a flagship model built for agentic engineering and long-horizon coding, capable of running up to eight hours on a single task.
π #ai
TestingCatalog
Zhipu AI launches open-source GLM-5.1 model for coding tasks
GLM-5.1 by Z.ai debuts as a coding-focused AI model supporting long autonomous tasks, now available for all GLM Coding Plan users.
β€4π3
Mythos grade intelligence might become available to users sooner than βmonthsβ.
> itβll probably be months before we use a model of this level of capability
> Uhm
Soon π
> itβll probably be months before we use a model of this level of capability
> Uhm
Soon π