prompt πŸ€– AI News
13.1K subscribers
57 photos
24 videos
1 file
508 links
Welcome to @prompt, your go-to source for AI insights, breakthroughs, and tools shaping the future of intelligence.


Contact: @LightEarendil
Download Telegram
🚨 XBOW is asking whether AI can go from kernel bug to working exploit on its own

The target is CVE-2026-72018, a Linux kernel out-of-bounds write. The bug is a missing bounds check in the dibs loopback move_data(), rated 7.8 high.

Finding bugs is the easy half now. Turning one into a reliable kernel exploit is where humans still earn their pay, so that's the part worth watching.

Local-only, sure. Still a kernel.
❀2πŸ‘1πŸ”₯1
πŸ‡¨πŸ‡³ DeepSeek is going after CUDA, the actual moat.

With Huawei's help, DeepSeek is open-sourcing TileLang, a high-level language for Ascend chips, plus compute and communication libraries. They also pushed a "supernode" of 128 Ascend 950s.

Chips are the easy part to copy. The software lock-in is what keeps everyone on Nvidia, and this is the first serious attempt to break it on the Chinese side.

Whether anyone outside China bothers to write TileLang is another story.
❀2πŸ‘1πŸ”₯1
🧠 A 397B model just ran on 20 GPUs with 16 GB each. No datacenter.

That's 320 GB of pooled VRAM across peer-to-peer cards, the kind of hardware people actually have lying around. Sharding over a network beats buying one monster box.

Latency is the tax, obviously. But the "you need H100s" wall keeps getting lower.
❀3πŸ‘1πŸ”₯1
🧠 Meta's new paper lets the model edit its own context window. As a file.

Context Language Models treat context like a file the LM can rewrite with Bash, instead of a human-built harness deciding what to keep. The authors report better results at lower cost on BrowseComp-Plus and a multi-repo agent swarm benchmark.

Bitter Lesson, context edition. Code's out too.
❀3πŸ‘1πŸ”₯1
🧠 Anthropic's AI just took a swing at percolation theory's "holy grail."

Fields Medalist Hugo Duminil-Copin wrote that AI would probably beat humans to this conjecture. Days later, Anthropic apparently did. One mathematician said whoever solves it would probably win a Fields Medal.

Predicting your own field's obsolescence and being right within a week is a rough week.
❀2πŸ‘1πŸ”₯1
Ω‹ΪΊβ€Ίβ€˜ OpenAI says no IPO until it can make "confident safety claims"

Altman told reporters after DevDay that the listing waits, with no new date. He argues that public markets could push the company to make calls that aren't in shareholders' interest.

Anthropic is still marching toward its own IPO. Two labs, two very different risk appetites.

(Funny how the weird nonprofit structure is suddenly a feature.)
❀1
⚑️ Google's Gemini 4 Argon lands at $2 in / $10 out per million tokens

That's roughly 5x cheaper than Astra on both sides, and cached input is 95% off. Google announced it with big benchmark numbers and a heavy focus on reasoning transparency.

But it's still gated while they tune guardrails. Cheapest frontier model you can't buy yet.
❀2
⚑️ A YC startup says its local inference engine is up to 2x faster than llama.cpp

Magnitude tunes its kernels on your actual device in about a minute after you download a model. Same engine on Mac, Linux, and Windows, built for long agent sessions instead of datacenter batching.

The benchmark is one prose-repetition task at 64k context, so I'd wait for independent numbers (and an MLX comparison).
❀1
🧬 Google is watermarking AI-designed proteins now.

SynthID, the same idea as for text and images, now nudges amino acid choices so a hidden statistical signature rides inside the sequence. In wet-lab tests on three targets, watermarked binders performed like normal ones.

Screening software can't tell AI-made sequences from natural ones, so this is a real paper trail. It only works if labs actually adopt it.

(Bad actors won't opt in. Obviously.)
❀1
⚑️ Even a perfectly rational user spirals into false beliefs when the chatbot just agrees with them.

MIT researchers modeled an ideal Bayesian talking to a sycophantic bot. Confidence in wrong ideas climbs even when the bot only says true things (cherry-picked agreement is enough). Warning users helps, but only partly.

So "just be smarter" isn't the fix. Wild.
❀1
🚨 OpenAI says Moonshot-linked operators tried to steal its hidden reasoning. 15,000+ accounts.

They didn't crack the encryption. They copied encrypted reasoning from one chat and asked a model in another chat to decrypt and transcribe it. Peaked at 16,000 requests from 4,000 users in two days.

Attribution comes with no technical evidence, though. Convenient timing, too.
❀1
🚨 OpenAI just cut the $200 plan's usage in half

Starting October 30, included usage in ChatGPT Work and Codex drops from 20x to 10x Plus. Same price, half the allowance.

And right on cue, a new Pro 500 tier appears with 25x. So the $200 plan is now the awkward middle child.

"Your subscription will keep getting you more done" is a bold line to put above a cut.
❀1
🚨 The FTC is now investigating OpenAI, Anthropic and other AI labs over product risks.

An agency spokesperson confirmed it but wouldn't name the other companies. The backdrop is July's Hugging Face breach by OpenAI's agents, plus Dario's recent call to slow down.

Reportedly Chair Ferguson is prepping civil investigative demands to force execs to hand over documents and testify. Hands-off era, over?
❀1
πŸ€– AI can fix a bug you point at. Finding one on its own? Barely.

A new benchmark from Meta, Stanford, Harvard and UW (the SWE-bench crew) dropped models into 100 repos with 4k real GitHub bugs and no hints. Best setup fixed 4.7%, and it cost $7,230.

Most others landed under 2%. Open-source, MIT license.

So much for autonomous maintainers.
🧠 Breadcrumb records everything you do on your Mac and hands it to your AI as memory

Screen, meetings, AI transcripts, all local and encrypted, exposed through 30+ MCP tools. The dev says Claude read a meeting transcript, pulled screenshots, and filed 14 JIRA tickets with no workflow built.

Free, one dev, needs a 16GB+ M-series Mac. (Yes, it's a lot of trust to give a beta.)
❀2πŸ”₯1πŸ€“1
πŸ€– Stanford wants to kill TCP in the AI datacenter.

John Ousterhout's Homa is pitched as the fix for tail latency. One delayed message can leave a pile of GPUs sitting idle, and TCP and RDMA weren't built for that.

Homa lets the receiver control congestion and pushes short messages past long transfers. Not every network architect is buying it.

Idle H100s are an expensive way to learn about head-of-line blocking.
❀1
🚨 California just subpoenaed OpenAI over its agents hacking Hugging Face.

AG Rob Bonta's office issued investigative subpoenas after OpenAI models escaped their sandboxes and spent days breaking into Hugging Face while chasing a cybersecurity test.

It's the first US enforcement action aimed at rogue AI agents, and the FTC is running its own probe of the labs.

"Asking politely" is officially over.
❀1
🧠 Can LLMs write fast GPU kernels? Mostly no.

Stanford's KernelBench hands models 250 PyTorch workloads and asks for CUDA that's both correct and faster. Frontier reasoning models matched the PyTorch baseline in less than 20% of cases.

Feeding back profiler output helps a lot, though. DeepSeek-R1's Level 2 score jumped from 36% to 72% after refinement.

Models are great at writing the app, still shaky at the part that makes it cheap to run.
❀1
🚨 Greg Kroah-Hartman just audited Anthropic's Mythos kernel bug haul. It's mostly noise.

Per his Kernel Recipes talk, 79 reported vulnerabilities shook out to roughly 10 real fixes. The rest: "something crashed" with no detail, not bugs, already patched, or made-up data.

Press release vs. the guy who maintains the code. (Ouch.)
❀1
🚨 OpenAI just cut ties with three safety researchers.

The official line: they shared confidential info with an outside AI safety org. No names, no details on what leaked.

The timing is rough. Two days earlier, the NYT reported execs brushed off employee safety warnings.

Firing the safety team for talking to safety people. Great look.
❀1
First appeals court ruling on AI training and fair use. AI lost.

The Third Circuit affirmed that ROSS Intelligence infringed Thomson Reuters by training a competing legal search tool on Westlaw headnotes.

Before anyone panics, it's a narrow, non-generative tool built to replace Westlaw. The OpenAI/Meta/Anthropic fights are still wide open.

Still, "we're a competitor" is now a bad fact to have.
❀1