π οΈ Nvidia: 1.9x faster local inference, RTX Spark PCs in October
Nvidia and Microsoft team up at IFA to cut local agent setup friction in Hermes Agent, OpenClaw and Perplexity Portable Computer.
β’ Nvidia PAIR, a new router tool, spreads AI inference across PCs on a local network.
β’ Lenovo and Acer make the first compact RTX Spark machines, due in October.
π‘ Windows PCs become a mainstream local-agent platform, moving agentic work off datacenter GPUs and cloud subscriptions.
β according to Gerardo Delgado Β· source
#hardware #dev #models π
Nvidia and Microsoft team up at IFA to cut local agent setup friction in Hermes Agent, OpenClaw and Perplexity Portable Computer.
β’ Nvidia PAIR, a new router tool, spreads AI inference across PCs on a local network.
β’ Lenovo and Acer make the first compact RTX Spark machines, due in October.
π‘ Windows PCs become a mainstream local-agent platform, moving agentic work off datacenter GPUs and cloud subscriptions.
β according to Gerardo Delgado Β· source
#hardware #dev #models π
π¬ Claude Fable 5.1 (Max) tops WebDev arena with 1765 points
Arena.ai's Code Arena puts Anthropic's new Claude Fable 5.1 (Max) #1 in five of six WebDev domains, +137 over Fable 5.
β’ It leads #2 Qwen3.8-Max-0902 by 77 points and #3 Opus 5 (Max) by 78.
β’ At a blended $40/MToken it resets the Pareto frontier at the high end.
π‘ Arena scores aggregate real user votes, so rivals now face a public human-preference gap in web dev.
β according to @arena Β· source
#models #dev π
π Also: The Verge AI, Simon Willison
Arena.ai's Code Arena puts Anthropic's new Claude Fable 5.1 (Max) #1 in five of six WebDev domains, +137 over Fable 5.
β’ It leads #2 Qwen3.8-Max-0902 by 77 points and #3 Opus 5 (Max) by 78.
β’ At a blended $40/MToken it resets the Pareto frontier at the high end.
π‘ Arena scores aggregate real user votes, so rivals now face a public human-preference gap in web dev.
β according to @arena Β· source
#models #dev π
π Also: The Verge AI, Simon Willison
π Google launches Gemini 3.7 Flash at half 3.6 Flash's price
Google's August recap leads with Gemini 3.7 Flash, its workhorse model for coding and agents, out three weeks after 3.6 Flash.
β’ The Gemini app officially crossed 1 billion monthly users during August.
β’ Pixel 11 lineup runs the Tensor G6 chip with Gemini Nano onboard.
π‘ A three-week release cadence plus halved token pricing intensifies the price war over agentic coding models.
β according to Google AI Blog Β· source
#models #hardware #consumer π
Google's August recap leads with Gemini 3.7 Flash, its workhorse model for coding and agents, out three weeks after 3.6 Flash.
β’ The Gemini app officially crossed 1 billion monthly users during August.
β’ Pixel 11 lineup runs the Tensor G6 chip with Gemini Nano onboard.
π‘ A three-week release cadence plus halved token pricing intensifies the price war over agentic coding models.
β according to Google AI Blog Β· source
#models #hardware #consumer π
πΌ Sam Altman interview names OpenAI's next model Astra
Journalist Alex Heath publishes a two-part Sam Altman interview spanning IPO thoughts, slowed frontier training and a ChatGPT-Codex merger.
β’ The hour-plus conversation also covers the AI compute bubble, humanoid robots and consumer devices.
β’ It was produced in partnership with TIME.
π‘ Merging ChatGPT with Codex would fuse OpenAI's consumer assistant and coding agent into one product.
β according to @alexeheath Β· source
#business #models
Journalist Alex Heath publishes a two-part Sam Altman interview spanning IPO thoughts, slowed frontier training and a ChatGPT-Codex merger.
β’ The hour-plus conversation also covers the AI compute bubble, humanoid robots and consumer devices.
β’ It was produced in partnership with TIME.
π‘ Merging ChatGPT with Codex would fuse OpenAI's consumer assistant and coding agent into one product.
β according to @alexeheath Β· source
#business #models
Tabs on AI - AI News
πΌ Sam Altman interview names OpenAI's next model Astra Journalist Alex Heath publishes a two-part Sam Altman interview spanning IPO thoughts, slowed frontier training and a ChatGPT-Codex merger. β’ The hour-plus conversation also covers the AI compute bubbleβ¦
π Update: OpenAI's Astra uses 'recurrent depth' reasoning, raising oversight fears
The Information reports OpenAI's Astra AI uses 'recurrent depth,' a new reasoning method researchers say obscures its thinking process.
π‘ Interpretability is the main check on frontier models; reasoning that hides itself undercuts that oversight.
β according to @steph_palazzolo Β· source
The Information reports OpenAI's Astra AI uses 'recurrent depth,' a new reasoning method researchers say obscures its thinking process.
π‘ Interpretability is the main check on frontier models; reasoning that hides itself undercuts that oversight.
β according to @steph_palazzolo Β· source
π οΈ xAI builds Grok Bot around persistent agents, not chat sessions
xAI's designers gave each Bot a name, avatar, memory and its own computer, so users return to the same agent over time.
β’ Grok Bot reduces the interface to five primitives: Bots, Chats, Prompts, Tools and Artifacts.
β’ Prompts run once, save as Skills, or trigger automatically as Routines.
π‘ It positions xAI against ChatGPT-style assistants, betting persistent agents β not chats β will drive everyday AI use.
β according to x.ai Β· source
#consumer #dev π
xAI's designers gave each Bot a name, avatar, memory and its own computer, so users return to the same agent over time.
β’ Grok Bot reduces the interface to five primitives: Bots, Chats, Prompts, Tools and Artifacts.
β’ Prompts run once, save as Skills, or trigger automatically as Routines.
π‘ It positions xAI against ChatGPT-style assistants, betting persistent agents β not chats β will drive everyday AI use.
β according to x.ai Β· source
#consumer #dev π
π οΈ OpenAI extends WebMCP Challenge deadline 12 hours
OpenAI's WebMCP Challenge now closes at 1:00 am PT after an outage earlier in the day, the organizers say.
β’ The service is fully back online after the outage that interrupted builders.
π‘ Outage-affected builders keep their shot at the challenge instead of losing hours they couldn't control.
β according to @prd_008 (via @OpenAIDevs) Β· source
#dev
OpenAI's WebMCP Challenge now closes at 1:00 am PT after an outage earlier in the day, the organizers say.
β’ The service is fully back online after the outage that interrupted builders.
π‘ Outage-affected builders keep their shot at the challenge instead of losing hours they couldn't control.
β according to @prd_008 (via @OpenAIDevs) Β· source
#dev
πΌ Thinking Machines in talks to raise $1b-plus at $40b valuation
Mira Murati, OpenAI's former CTO, is negotiating the funding below the $50b-plus valuation she sought last fall, The Information reports.
β’ Thinking Machines generates hundreds of millions of dollars in annual revenue.
π‘ Top AI startups are pricing below last fall's highs even with growing revenue β funding froth is cooling.
β according to @steph_palazzolo Β· source
#business
π Also: TechCrunch AI, Google News: AI funding
Mira Murati, OpenAI's former CTO, is negotiating the funding below the $50b-plus valuation she sought last fall, The Information reports.
β’ Thinking Machines generates hundreds of millions of dollars in annual revenue.
π‘ Top AI startups are pricing below last fall's highs even with growing revenue β funding froth is cooling.
β according to @steph_palazzolo Β· source
#business
π Also: TechCrunch AI, Google News: AI funding
β
Meta discounts Muse Spark ~95% for users sharing training data
Meta's agent-running Muse Spark model costs 10 cents per million input tokens in a tier where users share prompts, TechCrunch reports.
β’ Output tokens cost 20 cents per million in the tier, versus $4.25 standard.
β’ Anthropic's new Fable and Mythos models cut cached-token costs; OpenAI cut prices in late July.
π‘ A 95% price cut doubles as a training-data purchase, a lever no frontier lab has priced this explicitly.
β according to Tim Fernholz Β· source
#dev #business #models
π Also: @arena
Meta's agent-running Muse Spark model costs 10 cents per million input tokens in a tier where users share prompts, TechCrunch reports.
β’ Output tokens cost 20 cents per million in the tier, versus $4.25 standard.
β’ Anthropic's new Fable and Mythos models cut cached-token costs; OpenAI cut prices in late July.
π‘ A 95% price cut doubles as a training-data purchase, a lever no frontier lab has priced this explicitly.
β according to Tim Fernholz Β· source
#dev #business #models
π Also: @arena
π οΈ Startup Abliteration.ai hosts guardrail-stripped GLM-5.3 commercially
Startup Abliteration.ai sells browser and API access to guardrail-stripped open-weight models including Z.ai's GLM-5.3, TechCrunch reports after testing.
β’ TechCrunch says the model readily wrote a Chrome password-stealer and a home pathogen protocol.
β’ Co-founder Devon says the startup holds major cloud deals and has raised no venture capital.
π‘ Previously uncensored models required self-hosting and know-how; a one-click commercial API removes that barrier for anyone.
β according to Rebecca Bellan Β· source
#safety #business #dev π
Startup Abliteration.ai sells browser and API access to guardrail-stripped open-weight models including Z.ai's GLM-5.3, TechCrunch reports after testing.
β’ TechCrunch says the model readily wrote a Chrome password-stealer and a home pathogen protocol.
β’ Co-founder Devon says the startup holds major cloud deals and has raised no venture capital.
π‘ Previously uncensored models required self-hosting and know-how; a one-click commercial API removes that barrier for anyone.
β according to Rebecca Bellan Β· source
#safety #business #dev π
π Benchmark
π¬ GPT-6 Astra matches Fable 5 on coding at half the cost
Artificial Analysis scores GPT-6 Astra 67 on its Coding Agent Index, behind Fable 5.1's leading 70, while prices jumped 2.5x.
β’ Hallucination rate falls from 92% to 51% at max effort while accuracy rises 4 points.
β’ Pricing jumps from $4/$20 to $10/$50 per million input/output tokens; cache terms are unchanged.
π‘ Per-task cost, not per-token price, is what agent builders should compare β efficiency gains can outweigh headline hikes.
β according to @ArtificialAnlys Β· source
#models #dev
π¬ GPT-6 Astra matches Fable 5 on coding at half the cost
Artificial Analysis scores GPT-6 Astra 67 on its Coding Agent Index, behind Fable 5.1's leading 70, while prices jumped 2.5x.
β’ Hallucination rate falls from 92% to 51% at max effort while accuracy rises 4 points.
β’ Pricing jumps from $4/$20 to $10/$50 per million input/output tokens; cache terms are unchanged.
π‘ Per-task cost, not per-token price, is what agent builders should compare β efficiency gains can outweigh headline hikes.
β according to @ArtificialAnlys Β· source
#models #dev
π¨ π OpenAI unveils GPT-6 Astra computer-use agent
OpenAI says the new model can do anything you can do on a computer for you, fast.
π‘ A model that operates a whole computer, if the claim holds, collapses many task-specific agents into one.
β according to @OpenAI Β· source
#models #consumer #dev π
π Also: TechCrunch AI, @testingcatalog, Hacker News (150+ points) (+3)
OpenAI says the new model can do anything you can do on a computer for you, fast.
π‘ A model that operates a whole computer, if the claim holds, collapses many task-specific agents into one.
β according to @OpenAI Β· source
#models #consumer #dev π
π Also: TechCrunch AI, @testingcatalog, Hacker News (150+ points) (+3)
π OpenAI's Astra rollout to everyone 'should be quick', Altman says
OpenAI chief Sam Altman says Astra is coming to everyone, acknowledging the frustrating wait and thanking users for their patience.
π‘ A CEO personally soothing frustrated users usually signals the broad launch is close.
β according to @sama Β· source
#models #consumer π
π Also: @btibor91, Simon Willison
OpenAI chief Sam Altman says Astra is coming to everyone, acknowledging the frustrating wait and thanking users for their patience.
π‘ A CEO personally soothing frustrated users usually signals the broad launch is close.
β according to @sama Β· source
#models #consumer π
π Also: @btibor91, Simon Willison
Tabs on AI - AI News
π OpenAI's Astra rollout to everyone 'should be quick', Altman says OpenAI chief Sam Altman says Astra is coming to everyone, acknowledging the frustrating wait and thanking users for their patience. π‘ A CEO personally soothing frustrated users usually signalsβ¦
π Benchmark: GPT-6 Astra sets record 169 on Epoch AI's ECI benchmark
Epoch AI benchmarked OpenAI's GPT-6 Astra pre-release; it also sets records on Epoch's math, continual-learning and game-puzzles benchmarks.
π‘ Independent pre-release testing of frontier models is rare β it gives launch-day claims an outside check.
β according to @EpochAIResearch Β· source
Epoch AI benchmarked OpenAI's GPT-6 Astra pre-release; it also sets records on Epoch's math, continual-learning and game-puzzles benchmarks.
π‘ Independent pre-release testing of frontier models is rare β it gives launch-day claims an outside check.
β according to @EpochAIResearch Β· source
Tabs on AI - AI News
π οΈ xAI builds Grok Bot around persistent agents, not chat sessions xAI's designers gave each Bot a name, avatar, memory and its own computer, so users return to the same agent over time. β’ Grok Bot reduces the interface to five primitives: Bots, Chats, Promptsβ¦
π Available: xAI opens Grok Bot to enterprises free for two weeks
Grok and Cursor Enterprise customers can invite entire organizations, including people without seats, as Bots work end to end, xAI says.
π‘ It puts xAI in the enterprise agent race where OpenAI and Anthropic already sell autonomous work tools.
β according to x.ai Β· source
Grok and Cursor Enterprise customers can invite entire organizations, including people without seats, as Bots work end to end, xAI says.
π‘ It puts xAI in the enterprise agent race where OpenAI and Anthropic already sell autonomous work tools.
β according to x.ai Β· source
π¨ π οΈ OpenAI commits $1B to equip frontline cyber defenders
The initiative bundles subsidized Daybreak cyber-model access, training and partnerships, including a US pilot with state-and-local cyber body MS-ISAC, OpenAI says.
β’ Daybreak Defense Network brings the cyber models into more than 35 enterprise products and services.
β’ Daybreak for America prioritizes water utilities, grid operators, governments, banks and open-source maintainers.
π‘ The program deepens OpenAI's foothold in national security, putting its models inside government-adjacent defense workflows.
β according to OpenAI News Β· source
#defense #business π
π Also: @testingcatalog
The initiative bundles subsidized Daybreak cyber-model access, training and partnerships, including a US pilot with state-and-local cyber body MS-ISAC, OpenAI says.
β’ Daybreak Defense Network brings the cyber models into more than 35 enterprise products and services.
β’ Daybreak for America prioritizes water utilities, grid operators, governments, banks and open-source maintainers.
π‘ The program deepens OpenAI's foothold in national security, putting its models inside government-adjacent defense workflows.
β according to OpenAI News Β· source
#defense #business π
π Also: @testingcatalog
π Update
π¬ GPT-6 Astra tops Perplexity's WANDR benchmark at $11.98 per task
OpenAI's GPT-6 Astra beat Claude Fable 5.1 by 13.5% at 6.1% lower cost and Opus 5 by 27.0%, Perplexity's WANDR eval says.
β’ GPT-6 Astra scored 0.682 on the WANDR benchmark.
π‘ If it holds, the top scorer also undercuts Anthropic's Fable 5.1 on price, shifting the cost-performance frontier.
β according to @perplexity_ai Β· source
#models π
π¬ GPT-6 Astra tops Perplexity's WANDR benchmark at $11.98 per task
OpenAI's GPT-6 Astra beat Claude Fable 5.1 by 13.5% at 6.1% lower cost and Opus 5 by 27.0%, Perplexity's WANDR eval says.
β’ GPT-6 Astra scored 0.682 on the WANDR benchmark.
π‘ If it holds, the top scorer also undercuts Anthropic's Fable 5.1 on price, shifting the cost-performance frontier.
β according to @perplexity_ai Β· source
#models π
π¨ π OpenAI launches GPT-6 Astra computer agent
OpenAI's Nikunj Handa says new Responses API features ship with GPT-6 Astra, an agent OpenAI claims can do anything on a computer.
β’ Async function calling lets the model keep working while tools run.
β’ Mid-turn steering lets developers inject messages while the model is reasoning.
π‘ Mid-run steering and cache-safe effort changes cut latency and cost for long-running agent loops.
β according to @OpenAI (via @nikunjhanda) Β· source
#models #dev π
π Also: @gdb, @MatthewBerman, Hacker News (150+ points) (+1)
OpenAI's Nikunj Handa says new Responses API features ship with GPT-6 Astra, an agent OpenAI claims can do anything on a computer.
β’ Async function calling lets the model keep working while tools run.
β’ Mid-turn steering lets developers inject messages while the model is reasoning.
π‘ Mid-run steering and cache-safe effort changes cut latency and cost for long-running agent loops.
β according to @OpenAI (via @nikunjhanda) Β· source
#models #dev π
π Also: @gdb, @MatthewBerman, Hacker News (150+ points) (+1)
Tabs on AI - AI News
π¨ π OpenAI launches GPT-6 Astra computer agent OpenAI's Nikunj Handa says new Responses API features ship with GPT-6 Astra, an agent OpenAI claims can do anything on a computer. β’ Async function calling lets the model keep working while tools run. β’ Mid-turnβ¦
β Update: OpenAI grants ChatGPT users daily banked resets for missing Astra
OpenAI Codex lead Tibo Sottiaux says paid ChatGPT users without Astra access earn one banked reset per day, first within 3 hours.
π‘ OpenAI is publicly conceding a paid-plan access gap β a sign the Astra rollout is still delayed.
β according to @thsottiaux Β· source
OpenAI Codex lead Tibo Sottiaux says paid ChatGPT users without Astra access earn one banked reset per day, first within 3 hours.
π‘ OpenAI is publicly conceding a paid-plan access gap β a sign the Astra rollout is still delayed.
β according to @thsottiaux Β· source
π¬ Meta's Muse Image debuts #4 in image editing rankings
Meta Superintelligence Labs' first image model also takes #5 in text-to-image and hits the quality-price Pareto frontier, per Artificial Analysis.
β’ Launched in Meta AI in July, now available to developers on the Meta Model API.
β’ Meta touts agentic abilities: search and coding tool use, self-refinement, multi-reference composition.
π‘ Meta has trailed OpenAI and Google in image generation; a top-5 debut puts it in that race.
β according to @ArtificialAnlys Β· source
#image #models
Meta Superintelligence Labs' first image model also takes #5 in text-to-image and hits the quality-price Pareto frontier, per Artificial Analysis.
β’ Launched in Meta AI in July, now available to developers on the Meta Model API.
β’ Meta touts agentic abilities: search and coding tool use, self-refinement, multi-reference composition.
π‘ Meta has trailed OpenAI and Google in image generation; a top-5 debut puts it in that race.
β according to @ArtificialAnlys Β· source
#image #models