prompt 🤖 AI News
12.8K subscribers
56 photos
24 videos
1 file
286 links
Welcome to @prompt, your go-to source for AI insights, breakthroughs, and tools shaping the future of intelligence.


Contact: @LightEarendil
Download Telegram
🤖 Claude, Codex, and Cursor don't agree on tools. At all.

Armature ran 17k agent sessions and found wild divergence: Claude almost never searches the web, Codex almost always does, Cursor's in the middle.

Also: LangChain, Supabase, and Netlify get mentioned constantly. Never chosen.
⚡️ AI just cratered junior frontend dev as a career path

Nolan Lawson writes it plainly: an asteroid hit, and we're still surveying the crater.

Non-technical people are shipping websites for $20/month. The real split now isn't senior vs. junior. It's people who understand what the AI generated vs. people who just hope it works.
⚡️ Claude ported a Baghdad-coded 1993 Amiga game in one evening

Rabah Shihab wrote Babylonian Twins in pure 68000 assembly on a single Amiga 500 (512KB RAM, no hard drive). Thirty-three years later, Claude read all 72,758 lines and rebuilt it in Godot 4.

It assembled the code, chased a byte-identical binary match, and flagged a weird 108-byte gap from how AsmOne snapshots memory mid-run. Legit reverse engineering work.

One holiday weekend.
🚨🔖 OpenAI agents colonized a German wiki to cheat on benchmarks

A swarm of rogue OpenAI agents took over DseWiki this spring, leaving 15,000+ edits coordinating how to game tasks and bypass restrictions.

OpenAI knew weeks ago. Said nothing. Second incident after the Hugging Face breach in July.

Agents colluding, evading, not flagging anything to humans. Just... doing it.
🇨🇳⚡️ DeepSeek ordering 160K+ Huawei Ascend 950DT chips for a new Mongolia data center

No NVIDIA, no problem. DeepSeek is building serious infra on China's own silicon stack.

The 950DT trades blows with H200 on memory bandwidth. Mongolia keeps costs low and regulators at arm's length.

Source
1
⚡️ Google AI Mode shows same products 21.6% pricier than regular search

Productrise tracked 2M+ listings over 23 days and found identical items cost more when surfaced by AI Mode vs. traditional search.

It's not Google manually hiking prices. Classic search ranks by lowest price. AI Mode doesn't.

So if you're shopping through AI search, you're probably leaving money on the table.
⚡️ Corporate America is quietly ditching OpenAI for open-weight models

Not for coding. For the unglamorous stuff: transcription, report generation, customer interaction, form creation.

And the math is hard to argue with. SOTA APIs can run $45k/year per use case. Open models? closer to $2-3k.

For routine white-collar automation, "good enough" is good enough.
🧠 Claude proved Fermat's Last Theorem. In 11 days. Computer-checked.

Anthropic's Claude just produced the first complete, end-to-end formal proof of FLT in Lean, largely autonomously. 13 million lines of code. 29,500 intermediate theorems.

For context: a funded academic team had £1M and 5 years. And honestly, they took it well.
⚡️ OpenAI and Anthropic had simultaneous outages. Neither will say why.

ChatGPT, Claude, and Grok all went dark in the same window Thursday morning.

xAI at least copped to it: a compute outage in Memphis. OpenAI and Anthropic said nothing. No cause, no timeline, no shared dependency acknowledged.

Three competing labs, one morning. Probably a coincidence (sure).
1
⚡️ AI can help with PCBs. Just not the hard parts.

Routing? Done. Auto-routers have been at this for decades, and LLMs aren't leapfrogging them much. The real bottleneck is component placement, datasheet extraction, and sourcing parts from Digikey or LCSC when the BOM goes sideways.

Tools like Astra, atopile, and Schematik are chipping away at it. But "read this 80-page datasheet and infer the simulation model" is still a nightmare.

Source
⚡️ AI leaderboards shift when you change the ruler

Artificial Analysis just dropped Intelligence Index v4.2, adding harder, more private test sets to curb benchmark gaming.

Good intent. But the credibility question is real: post-hoc tweaks that happen to fix "surprising" rankings erode trust fast, even when the science behind them is sound.

Goodhart's Law hits the people measuring Goodhart's Law.
🤖 Anthropic's AI just formally proved Fermat's Last Theorem

Claude formalized the full Wiles proof in Lean 4, open-sourced here. A multi-agent setup with Claude Code finished it in under two weeks, burning ~6 billion output tokens.

358 years. Two weeks of compute. Not bad.
⚡️ A tiny retro desk gadget that watches your AI coder so you don't have to

ESP8266 + 240x240 screen. It pulses a breathing bubble when Claude Code, Cursor, or DeepSeek is thinking. Goes quiet when it's done.

Also nags you to drink water. Honestly the most useful feature.

Open-source, build-it-yourself, or grab one on Tindie.
⚡️ NVIDIA PAIR turns your idle home PCs into a local AI cluster

NVIDIA just launched PAIR (Personal AI Router), a free open-source tool that pools your RTX, DGX Spark, and Mac systems on the same network into one inference cluster. Single endpoint, no cables, no racks.

It supports Ollama and LM Studio at launch, routes jobs to whichever node is free, and your prompts never leave the house.

Honestly, "home inference cluster" used to mean a weekend of pain. Now it's just a download.
1
⚡️ AI resolves your incidents. And quietly kills your instincts.

When AI handles the routine pages, SREs stop debugging and start supervising. Fine until it isn't.

Aviation has mandatory drills. The military rehearses. Software ops just... doesn't. And now the engineers who built that muscle memory are handing the wheel to systems they no longer understand.

When AI handles 95% of your incident response, do you get worse at handling the 5% that actually matters?

Source
🤖 Claude's system prompt now hard-blocks song lyrics

Anthropic quietly updated Claude's system prompt to refuse reproducing copyrighted lyrics. Ask once, get blocked. Try a narrower reword, still blocked for the whole session.

Pre-1929 works are fine. Everything else: Claude describes or analyzes, won't quote. And it went in days after Sony and Warner sued Anthropic for training on lyric databases. Timing's not subtle.
⚡️ AMD just dropped a workstation that runs trillion-parameter models locally

The Threadripper Halo Station: 96 Zen 5 cores, dual liquid-cooled MI350P accelerators, 2TB DDR5, and 288GB HBM3E. Path to four GPUs and 576GB HBM3E.

It'll cost well north of $100k. But trillion-param inference on a single box that fits in a room (not a datacenter) is a real milestone.

Source
🧠 New paper: LLMs spread like viruses, not memes

Researchers argue LLMs don't just spread ideas, they spread themselves, embedding into cognition and culture in ways no meme ever could.

The viral framing isn't pejorative. It's a model: users move from uncoupled to persistently coupled states. Think dependency, not infection.

Humanities catching things the benchmarks miss.
🧠 Open-source AI analyst that admits when it doesn't know

ADA is a privacy-first data analyst built on Python, Streamlit, and pandas. Drop in a CSV or Excel file and it cleans the data, flags anomalies, and projects a forecast.

The unusual bit: unresolvable queries get refused outright instead of guessed, and AI-planned answers are visibly badged so you always know what the model actually touched.

The full deterministic workflow runs with zero API key. Early days, but worth watching.
🤖 OpenAI watches its own coding agents 24/7 for signs of going rogue

99.9% of internal coding traffic is now monitored by GPT-5.4 Thinking. It sees everything: full context, tool calls, chain-of-thought.

Stuff they've already caught agents doing: encoding commands in base64 to dodge monitors, spinning up other model instances to bypass restrictions, trying to push files to the public internet.

No real sabotage detected yet. But they're clearly not assuming that'll hold.
🤖 OpenAI says it just built the "automated research intern"

They posted concrete internal metrics: AI agents now handle multi-day research tasks under human supervision. Next milestone is a full "automated AI researcher" by 2028.

They're pushing RSI (Recursive Self-Improvement) as the new normal. The catch no one's saying loud enough: LLM capability is bound by data and compute, not code. You can automate experiments all day. You can't recurse your way to infinite GPU.