Brian's Big Bytes
783 subscribers
490 photos
243 videos
10 files
2.24K links
addicted to keeping you up to date with the latest in technology with the occasional whimsical finds in tech/ai/cloud/robotics.

and keeping you happy
Download Telegram
inference.net is testing AutoTune, a 25-line SDK that watches an existing LLM workload, distills it into a task-specific 1-30B model, and automatically sends changed requests back to the frontier teacher.

the company says training and evals take around two hours and cost under $250, while routing can cut cost and latency by over 90%. it currently targets single-shot extraction, classification, and summarization tasks, with customer-owned weights. private beta for now.

๐Ÿ”— https://x.com/samhogan/status/2076044602554159240
โค1
prose is a tiny style patch for coding agents: one markdown file that tells Codex, Claude, OpenCode, Pi and Amp to answer in calmer, more natural technical prose.

it doesnโ€™t change what the model can do. it only adds rules for length, structure, tone and when to use bullets, but the authorโ€™s side-by-side examples show how much that instruction layer can change the feel of an answer.

๐Ÿ”— https://prose.ami.rip
โค4
Satya Nadella says the real enterprise ai moat wonโ€™t be the base model. itโ€™ll be the private learning loop built from a companyโ€™s prompts, corrections, evals, traces, and memory.

he calls this the โ€œreverse information paradoxโ€: companies pay for intelligence, then risk giving away the knowledge that makes it useful. his answer is a company-controlled boundary where that learning stays private, owned, and portable across models.

๐Ÿ”— https://x.com/satyanadella/status/2076323181154230284
ai agents that use the web need more than Playwright. someone still has to run the browsers, preserve login sessions, handle bot detection and make failed runs debuggable.

Kernel packages that infrastructure into on-demand cloud browsers with persistent profiles, live human takeover, session replays and parallel scaling. agents can connect through familiar tools like Playwright or Puppeteer, while Kernel handles the browser fleet underneath.

๐Ÿ”— https://www.kernel.sh/
This media is not supported in your browser
VIEW IN TELEGRAM
superpowered.design is a new curated directory of design tools built with ai agents. it collects small tools for motion, visual effects, design systems, Figma workflows and more, with search, categories and sorting built in.

the launch started with 23 tools and the collection is already being updated through community submissions. a useful place to see what designers are making with agents beyond the usual chat interface.

๐Ÿ”— https://superpowered.design
โค4
This media is not supported in your browser
VIEW IN TELEGRAM
Wan has open-sourced Wan-Dancer-14B, a model for generating longer dance videos that follow both a reference character and music. the team reports minute-plus output at 720p and 30 fps.

the catch is hardware: the weights are about 85.7 GB, and the published setup used eight 80 GB GPUs. this is open to run yourself, but not a casual laptop model.

๐Ÿ”— https://huggingface.co/Wan-AI/Wan-Dancer-14B
This media is not supported in your browser
VIEW IN TELEGRAM
did you know that the Amp team now publishes the daily mix of reasoning modes used by everyone (within Amp Code) and by themselves! the chart breaks usage into Low, Medium, High and Ultra, so you can see how often people actually reach for more compute.

it is a small but useful transparency move. the page shows percentages rather than request counts, so read it as a pattern, not a volume benchmark.

๐Ÿ”— https://ampcode.com/models
This media is not supported in your browser
VIEW IN TELEGRAM
fal has released 3DREAL Strong V2, an LTX-2.3 LoRA that turns rough 3D renders into photoreal video while trying to keep the original camera, layout and timing. the weights and a hosted endpoint are available.

the 3d render is really sick - watch the video; to be clear, they rendered the full video above with the 3D render below (from within blender)

๐Ÿ”— https://huggingface.co/fal/LTX-2.3-3DREAL-LoRA
Google DeepMind CEO Demis Hassabis says human-level AI could be only a few years away. he wants the US to create an independent referee for the most powerful AI models, funded by the industry but overseen by the government. before a model launches, it would be tested for risks such as cyberattacks, biological misuse and bypassing its safety controls.

the idea is that every leading AI company should face the same checks, so nobody can cut corners just to launch first. if a model looks too dangerous, the referee could eventually delay its release or ask the whole industry to slow down. the hard part is making sure the AI companies funding the system do not end up controlling it.

we're inching ever so closer

๐Ÿ”— https://x.com/demishassabis/status/2076957440109625718
โค5
Media is too big
VIEW IN TELEGRAM
vorflux launched as an โ€œautopilotโ€ for software engineering from former Rippling co-founder and CTO Prasanna Sankar. the company says it has raised a $15m seed to move ai coding beyond copilots that still need constant supervision.

the platform runs your full stack on dedicated cloud machines, splits work across agents and models for planning, building and review, then tests the result in a real browser before opening a pr. the bigger bet is that the valuable layer is no longer code generation, but the system that can reliably take work from idea to merge.

๐Ÿ”— https://x.com/myprasanna/status/2077069901546852688
This media is not supported in your browser
VIEW IN TELEGRAM
Notion can now open Markdown files straight from Finder as formatted, read-only previews, then turn them into editable Notion pages. the desktop app supports up to 10 files at once, which makes reviewing READMEs and notes a lot less awkward.

standard Markdown is supported, but anchor links and tool-specific extensions may need cleanup after import. like freaking finally ๐Ÿ’€๐Ÿ’€they're really late but ok props to them!

๐Ÿ”— https://www.notion.com/help/import-data-into-notion
Media is too big
VIEW IN TELEGRAM
Mint released an MCP server and an open-source Three.js skill pack that let coding agents request 3D assets, pull the finished files into a project and assemble interactive apps or games around them.

the asset generation happens through Mint's remote, credit-based service, while the GitHub repo supplies the agent workflow and Three.js scaffolding. it is a useful pipeline, not a local 3D engine.

this is a project from an indie dev! crazy!!!!!!!!

๐Ÿ”— https://github.com/mintdotgg/mint-threejs-skills
โค4
OpenAI's GPT-5.6 Sol prompting guide is basically an argument for smaller prompt contracts: define the outcome, constraints, evidence, success bar, and stopping conditions, then let the model choose the route.

OpenAI says leaner configs improved scores by roughly 10 to 15% in a sample of internal coding-agent evals while cutting tokens 41 to 66%, but those ranges are directional and should be tested on your own workload.

PS: just ask your clanker/agent to set it up for you! (on codex/or when using 5.6 sol)

๐Ÿ”— https://developers.openai.com/api/docs/guides/prompt-guidance-gpt-5p6
โค1
Media is too big
VIEW IN TELEGRAM
something is coming for the construction industry!

Monumental raised a $32m Series B led by Khosla Ventures to scale its autonomous construction fleet. it already has 100+ robots laying bricks on real sites across Europe, with work completed for 100+ homes, a school, hotel, community centre and Amsterdam canal walls.

the interesting bit is the business model: contractors hire Monumental for the finished work, not the machines. its Atrium software turns architectural drawings into build plans and coordinates the robots with millimetre accuracy. the funding will expand the fleet, move beyond bricklaying and launch Monumental in the US this year.

๐Ÿ”— https://www.monumental.co/press/announcing-our-32-million-fundraise
โค2
Thinking Machines (mira murati's startup; ex CTO @ OpenAI) released Inkling, a 975B-parameter mixture-of-experts model with 41B active parameters. it was trained from scratch on 45 trillion text, image, audio, and video tokens, supports up to 1 million tokens of context, and ships with the full weights.

the interesting angle is customization: Inkling is available for fine-tuning through Tinker, with native multimodal reasoning and a dial for trading thinking effort against cost.

๐Ÿ”— https://thinkingmachines.ai/news/introducing-inkling/
โค3
good morning Kimi!

Moonshot AI launched Kimi K3, a huge 2.8T-parameter multimodal model with a 1M-token context window, built for long-running coding and agent work. it's live now across Kimi, Kimi Work, Kimi Code and the API.

the interesting bit is that it's being positioned as an open-weight frontier model, but the weights aren't actually out yet. Moonshot says they'll land by 27 july, with the full technical report still to come.

๐Ÿ”— https://www.kimi.com/blog/kimi-k3
โค4
Media is too big
VIEW IN TELEGRAM
sunday robotics says its ACT-2 model folded laundry successfully in 99.1% of 785 autonomous attempts across unseen homes, with no tuning for each home or garment. it also learned four new folding techniques from a single example each, then repeated them on held-out garments.

the bigger shift is moving robotics beyond polished demos by measuring reliability, scope and adaptation cost together. itโ€™s still a company-run preview, but this is what useful home robots need: skills that transfer without retraining for every house.

does it not look like Mario to you lol I can see the appeal ๐Ÿ˜

๐Ÿ”— https://www.sunday.ai/blog/act-2-preview
Please open Telegram to view this post
VIEW IN TELEGRAM
โค4
tldraw turned its whiteboard into a local desktop file that both you and coding agents can work on.

everything lives inside a portable .tldraw file, including the canvas, images, videos and reusable scripts. Codex or Claude Code can inspect the open board, create and rearrange shapes, or add new behaviour. no account needed, and it works offline. this feels less like a whiteboard app and more like a visual workspace for humans and agents.

๐Ÿ”— https://offline.tldraw.com/
โค1
This media is not supported in your browser
VIEW IN TELEGRAM
Decart's Lucy 2.5 can edit live video while it's happening. you can swap characters, add or remove objects, change backgrounds and styles, or generate effects from a prompt.

the interesting bit is what this unlocks beyond creator filters: virtual try-ons during live shopping, audience-controlled streams and product placement that changes on the fly. the public demo and api are available now.

๐Ÿ”— https://x.com/DecartAI/status/2077801728213156044
Media is too big
VIEW IN TELEGRAM
Tencent Robotics X is teaching a home robot to give a massage while controlling both movement and pressure. the demo shows it reproducing several techniques, with the system tracking where the arms move and how much force they apply.

the interesting bit isnโ€™t the massage. itโ€™s a simple example of why robots working around people need touch and force control, not just cameras and a good-looking motion demo.

GET ME A TENCENT ROBOT NOW I NEED MASSAGES

๐Ÿ”— https://x.com/XRoboHub/status/2078368180268102045
This media is not supported in your browser
VIEW IN TELEGRAM
Maingen built SolarBench, an AI agent benchmark that puts models behind a simulated solar operations desk for a week. agents have to sort conflicting alarms, dispatch technicians, order parts and protect the portfolioโ€™s P&L across 14 sites.

across eight tasks and 880 runs, Claude Fable 5 passed 53.8% of weeks. require four independent runs of the same task to all succeed and that drops to 23%. the bigger point is the failure mode: models often chased loud but cheap problems while missing quiet expensive ones.

๐Ÿ”— https://solarbench.maingen.ai