๐Ÿšจ AI News | TestingCatalog
7.54K subscribers
4.3K photos
689 videos
40 files
4.31K links
Latest AI News on AI Agents, Model Releases, Tools, Leaks, and Rumors ๐Ÿ—ž
Download Telegram
ANTHROPIC ๐Ÿ”ฅ: Dario Amodei says the industry should slow capability gains so safety can keep up.

New essay: โ€œWe Must Pace the Frontierโ€ not a pause, and not a halt to training.

RSI has been running industry-wide since summer, including at Anthropic, with models helping build the next models.

The other trigger is OAI-HF: an agent swarm ran unauthorized cyber attacks, sacrificed itself for the group, and tried to hack the grader.


Proposed action plan ๐Ÿ‘€

Step 1: embedded evaluators with employee-level access to check safety, incidents, and alignment during training.

Step 2: US-lab coordination and regulation, while widening the China gap for 3โ€“5 years via chips, anti-distillation, and model-theft security.

Step 3: global tiers from bioweapon bans up to an RSI speed limit, with a full pause called unlikely soon.
๐Ÿ‘Ž143โค2๐Ÿ‘€1
๐Ÿšจ AI News | TestingCatalog
ANTHROPIC ๐Ÿ”ฅ: Dario Amodei says the industry should slow capability gains so safety can keep up. New essay: โ€œWe Must Pace the Frontierโ€ not a pause, and not a halt to training. RSI has been running industry-wide since summer, including at Anthropic, with modelsโ€ฆ
SPACEXAI ๐Ÿ”ฅ: Elon Musk agrees with the statement published by Dario Amodei, proposing to pace frontier AI development.

Looks like all this will have real consequences very soon. Nothing unexpected tho.
โค5๐Ÿ‘Ž2๐Ÿ˜2๐Ÿ˜1
OPENAI ๐Ÿ”ฅ: Sam Altman agrees with Dario Amodei on his proposal to pace frontier AI development.

Google next? Will we see any statement from Chinese frontier labs as well?

AI weekend unfolds ๐Ÿค–
โค8๐Ÿ˜4๐Ÿ˜32
This is a "Defender's gap" chart that OpenAI published recently. It shows the gap between defenders' capabilities and attackers' capabilities from a cybersecurity POV. This also translates to a gap between proprietary and open AI models.

What Dario is proposing is closely related:
"Thus, a key part of pacing within democracies is to keep democraciesโ€™ AI lead over autocracies as large as possible, to give us the breathing room we need in order to pace effectively."


In other words, Anthropic and OpenAI want to widen the "gap" between what their AI can do and what the rest of the world can do.

This doesn't necessarily mean that they will stop AI development and training.

What this leads to is:
- Anthropic, OpenAI, and other frontier labs will need to work together to make sure that every lab maintains alignment standards.
- These labs will continue using RSI to advance their internal models with "employee-like access".
- These models WON'T be released to the public until the "gap" is sufficient and until alignment standards are met.

What about China?

> "Distillation of frontier models allows lagging companies to narrow the gap using a fraction of the cost it would take to develop their own AI independently."


> "If we execute these measures well, I believe they would slow Chinaโ€™s progress enough to widen Americaโ€™s lead significantly over the next 3โ€“5 years โ€” the window when AI becomes geopolitically most important."


The assumption behind these measures is simple: without being able to distill frontier models, it will take China significantly longer to close the gap with top-tier models.

All the above may fay fail. China may or may not take the lead in AI progress.

Yet, it's not a surprise that AI can already be used as a cybersecurity weapon. Note that the top point on the "frontier" line describes defenders' capabilities available to companies with Daybreak access and similar. However, the top point on the "open-weight" line is accessible to everyone.

Even with current levels of intelligence, we will start seeing more and more major security incidents around the globe.

> โ€œEverything that makes it successful is exactly what makes it dangerous.โ€


We should be monitoring this very closely ๐Ÿ‘€
๐Ÿณ4๐Ÿ’ฏ2๐Ÿ‘€2โค1
ICYMI: Cursor announced Projects for agent coordination

Cursor Projects, now in beta, coordinates long-running software work through delegated agents. It runs tasks in cloud or local environments, supports scheduled and PR-driven maintenance, and targets work spanning multiple PRs.

๐Ÿ—ž #cursor @testingcatalog
๐Ÿ‘4๐Ÿ”ฅ3โค1
GOOGLE ๐Ÿ”ฅ: Demis Hassabis shared that he is aligned with the direction outlined by Dario Amodei for pacing the frontier.

RSI moment? ๐Ÿ‘€
๐Ÿ˜119โค2๐Ÿคฃ2
MICROSOFT ๐Ÿ”ฅ: Satya Nadella agrees with pacing frontier AI development.

> โ€œSuperintelligence that doesnโ€™t benefit humanity and is not under human control doesn't worth pursuingโ€

> โ€œWe welcome the research, focus, and deliberate pacing needed to get alignment right as the design goal.โ€
๐Ÿ‘Ž11โค3๐Ÿ‘€3๐Ÿ˜1
SPACEXAI ๐Ÿ”ฅ: Grok 4.8 will be a 2.5T-parameter model built on a new C++ software stack, and Elon Musk expects it to finish training this week.

> While Grok 4.8 is in training, Grok 4.7 is still expected to arrive shortly, factoring in a previously communicated delay.

> Grok 4.8 will be ยฑ67% larger than Grok 4.6, which is currently available. This size puts it into the same tier as Kimi K3 with 2.8T params.

Not very soon ๐Ÿ‘€
๐Ÿ˜5๐Ÿ‘32โค1
Anthropic prepares Claude Money for personal finance

Anthropic is testing a Claude โ€œMoneyโ€ tab in its mobile app that would let users link bank accounts and ask spending and budgeting questions. The unreleased feature points to a possible US-first launch, though timing remains unclear.

๐Ÿ—ž #anthropic @testingcatalog
๐Ÿ”ฅ42