Brian's Big Bytes
783 subscribers
491 photos
246 videos
10 files
2.24K links
addicted to keeping you up to date with the latest in technology with the occasional whimsical finds in tech/ai/cloud/robotics.

and keeping you happy
Download Telegram
This media is not supported in your browser
VIEW IN TELEGRAM
Dilum Sanjaya demoed a practical version of vibe coding for hardware ideas: generate a 3d gadget prototype in code, test the screen/buttons/interactions, then reuse parts of the logic when moving toward a real Raspberry Pi build.

the useful bit is the process, not the polish. he starts rough, refines the shape and materials later, adds sliders for tiny tweaks, and treats the generated model as a bridge between UI experiments and physical devices.

might be a bit slow on updates but they’re still coming!!!! (On a flight) ✈️

πŸ”— https://x.com/DilumSanjaya/status/2070912580122755484
Meta shared Brain2Qwerty v2, a non-invasive brain-to-text research system that decodes typed sentences from MEG brain recordings without a surgical implant. the jump is in words, not just characters: Meta says it trained on about 22k sentences from 9 volunteers and reached 61% average word accuracy, with its best participant at 78%.

important caveat: this is still lab research with bulky MEG hardware, not a consumer brain keyboard. but Meta is releasing the training code for v1 and v2, plus the v1 dataset, so other labs can poke at the path from brain signals to usable communication.

πŸ”— https://ai.meta.com/blog/brain2qwerty-brain-ai-human-communication/
Railway made a tiny fake meditation retreat for burnt-out developers, complete with a 45-second film, β€œdeploy code / touch grass,” and testimonials from engineers who have apparently deshrimped.

it’s marketing, but the joke lands because every cloud/devtools company is now selling speed while developers are quietly drowning in deploy anxiety, vibe-coded messes, and friday prod changes. Railway’s angle: shipping software should feel calmer, not more chaotic.

πŸ”— https://railway.com/peace
❀1
OpenClaw is now on iOS and Android. the useful bit is that its agents are no longer tied to a desktop or chat setup: you can run tasks, handle replies, and review approvals from your phone while the gateway stays under your control.

the iOS listing describes it as a local-first assistant paired to a private OpenClaw Gateway, with chat, Talk mode, approvals, sharing, and optional phone permissions. Android is live too, so mobile is now part of the core OpenClaw workflow.

πŸ”— https://x.com/openclaw/status/2071688039114342592
Extend is building document processing infrastructure for the part of AI agents that usually breaks first: messy PDFs, scans, tables, forms, handwriting, and document bundles.

the useful bit is that it's not just OCR. Extend gives teams parsing, extraction, classification, splitting, workflows, and evals, with Parse 2.0 turning complex documents into clean, agent-ready markdown instead of another pile of brittle text.

πŸ”— https://www.extend.ai/
Google is opening up more of its generative media stack to developers. Nano Banana 2 Lite is now generally available as its fastest, most cost-efficient Gemini image model, while Gemini Omni Flash is in public preview for video generation and conversational editing in Google AI Studio and the Gemini API.

the useful bit is the workflow: generate an image quickly, then pass it into Omni Flash to animate or keep editing through the Interactions API. it makes image-to-video feel more like a buildable product surface, not just a demo.

πŸ”— https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-omni-flash-nano-banana-2-lite/
This media is not supported in your browser
VIEW IN TELEGRAM
Anthropic launched Claude Sonnet 5, a cheaper Sonnet model meant for agent-style work: planning, using browsers and terminals, checking its own output, and finishing multi-step coding tasks with less hand-holding.

the notable bit is positioning: Anthropic says it gets close to Opus 4.8 on agentic tasks, is now the default for Free and Pro, and is on the API at intro pricing of $2 / $10 per million input/output tokens through Aug 31.

πŸ”— https://www.anthropic.com/news/claude-sonnet-5
❀3
Anthropic says the Commerce Department has lifted the export controls that forced it to pull Claude Fable 5 and Mythos 5 earlier this month. access is supposed to start coming back from July 1 U.S. time, after a strange 18-day pause where a model launch became a live national-security and compliance fight.

the bigger story is that frontier model access is now getting negotiated in public with Washington, not just shipped through product pages. for builders, the question is less "is the model good?" and more "who is actually allowed to use it, and under what rules?"

πŸ”— https://x.com/AnthropicAI/status/2072106151890809341
❀1
Etched is out of stealth with its first inference racks in customer validation: a successful A0 tapeout, $800m raised, and $1b+ in customer contracts. it says the first racks ship this summer.

the interesting part is that this is a full systems bet, not just a chip launch. Etched is pitching rack-scale inference clusters built around low-voltage compute and shared cluster memory to make frontier model inference faster, cheaper, and less power-hungry than today's GPU setups.

πŸ”— https://x.com/Etched/status/2071972062202343590
Media is too big
VIEW IN TELEGRAM
oasis launched OASIS 1, a smart ring for private dictation and touch control. the pitch is less fitness tracker and more input device: whisper to write, then tap or swipe on the ring to edit, scroll, pause music, or move around apps.

the useful angle is how narrow it starts. OASIS is selling the awkward middle ground where voice is faster than typing but speaking out loud feels too public. preorders are listed at $289, with shipping around Christmas 2026.

πŸ”— https://x.com/oasisdevices/status/2072033581241683988
❀1
very very sorry for the lack of news!!!!!! i’ve been spending too much time in person in San Francisco, and honestly it’s been jsfhifasjkndhgfq!!!!!!!!

there’s something about being here that’s hard to get online. even at afters, every single person you meet is building something, poking at some weird idea, or asking better questions than the internet usually gives you. that kind of curiosity is contagious. lime rides all around the city, the weather helps too.

17 hours flight to go!!!!!!!!!!!!! in the meantime, come check out this event where we'll try and talk about some of the cool shit that transpired!

speak more soon!
πŸ”— https://luma.com/olsz24cx
❀29
This media is not supported in your browser
VIEW IN TELEGRAM
dayflow launched an open source automatic work journal for Mac. it turns screen activity into a readable timeline of what you actually worked on, so you can reconstruct the day without timers, tags, or pretending your calendar was the truth.

the useful bit is the privacy model: recordings and the database stay on your Mac, and you can run analysis with local models or bring Gemini, ChatGPT, or Claude if you want better summaries. it’s MIT licensed, macOS 14+, and the repo is at ~6.5k stars.
πŸ”— https://x.com/jerryliu/status/2073116662602342734
❀2
This media is not supported in your browser
VIEW IN TELEGRAM
humanoids are usually framed as an AI problem, but Humanity's Last Machine is a good reminder that the body is the hard part too. it maps the actual hardware stack: skeleton, motors, reducers, screws, bearings, actuators, batteries, compute, sensors, tactile skin, and hands.

the most useful takeaway: costs probably compress more like cars than iPhones. actuators and hands are still the blocker, and the US-China split is less "who has better demos" than "who can iterate hardware, suppliers, and demand loops faster."

πŸ”— https://www.humanityslastmachine.com/
❀2
This media is not supported in your browser
VIEW IN TELEGRAM
Anthropic published a short history of Claude Code, and the interesting part is how long the core bet has been around: coding was the path to agents that can actually do work, not just answer questions.

the story traces it from an early VS Code assistant and internal clide / Claude CLI experiments to the february 2025 Claude Code research preview. the takeaway is that the product was less one big reveal than years of scaffolding, tool use, bash access, permissioning, and users teaching Anthropic what the workflow should become.

πŸ”— https://www.anthropic.com/features/making-of-claude-code
Media is too big
VIEW IN TELEGRAM
Anthropic published new interpretability research on what it calls Claude's "J-space": a small internal workspace where the model appears to hold concepts it can report, focus on, and use for multi-step reasoning, even when those thoughts never show up in the output.

important caveat: this is not proof that Claude has feelings or human-like consciousness. the practical part is a lens into some hidden internal state, which could help audit when a model notices tests, fabricates data, or pursues a goal it is not saying out loud.

πŸ”— https://www.anthropic.com/research/global-workspace
This media is not supported in your browser
VIEW IN TELEGRAM
grok's voice stack is getting a bigger cast: SpaceXAI added 21 new multilingual flagship voices to the Grok API, alongside upgrades to the original five.

the useful part is that these are not just app voices. they're available through the realtime Voice Agent API, Text to Speech API, and the new Voice Agent Builder, so developers can plug them into support agents, characters, ads, education, and other voice workflows.

πŸ”— https://x.ai/news/new-flagship-voices
❀1
This media is not supported in your browser
VIEW IN TELEGRAM
Meta introduced Muse Image and previewed Muse Video, the first media generation models from Meta Superintelligence Labs. the interesting bit is that Muse Image is more agentic than a normal image model: it can use search and code, self-refine, edit in place, and combine multiple reference images in one prompt.

Muse Image is available in Meta AI/meta.ai, Instagram Stories in the US, and WhatsApp in limited countries. Muse Video is still a preview, but Meta says it brings native audio support and is coming to creators and Meta AI.

πŸ”— https://ai.meta.com/blog/introducing-muse-image-muse-video-msl/
πŸ”— https://x.com/mattdeitke/status/2074556783583191432
❀1
Media is too big
VIEW IN TELEGRAM
Brainbase launched what it calls an AI agent cloud: infrastructure for deploying lots of agents across different models and tool harnesses, with the boring-but-important pieces handled for you: sandboxing, routing, evals, monitoring, and scaling.

the bigger idea is agents that provision more like compute. instead of one giant model doing everything, teams can spin up smaller purpose-built agents for workflows like PR review, incident triage, or customer requests, then route between models based on cost and performance.

πŸ”— https://x.com/BrainbaseHQ/status/2074530735911047608
❀1
OpenAI says GPT-5.6 Sol, Terra and Luna will launch publicly on Thursday, July 9, with preview access expanding globally now.

the bigger signal is access: this was previously a limited preview, and now the full model family is moving toward public availability instead of staying behind case-by-case approvals.

ARE... YOU... READY???????????????????????????/

πŸ”— https://x.com/openai/status/2074704958419792299
Cognition launched SWE-1.7, its new coding model for Devin, across web, desktop, and CLI. the practical bit: Cognition says it runs at 1000 tokens/sec via Cerebras, is free for paid users for the next month, and lands close to frontier models on its own FrontierCode eval while costing much less per task.

the interesting part is the training story: SWE-1.7 starts from Kimi K2.7, then Cognition pushes it with RL for longer software tasks, including self-compaction so it can summarize its work and keep going. the caveat is the headline benchmarks are mostly Cognition-run, so treat the numbers as launch data, not neutral third-party validation.

πŸ”— https://x.com/cognition/status/2074882968770728416
πŸ”— https://cognition.com/blog/swe-1-7
Media is too big
VIEW IN TELEGRAM
SpaceXAI launched Grok 4.5, a coding-and-agent model trained alongside Cursor. the practical pitch is less about chatbot vibes and more about engineering workflows: big codebases, long-running tasks, multi-repo work, and tool-heavy agent loops.

the launch claims 80 tokens/sec, roughly 2x token efficiency versus comparable models, and pricing at $2/M input and $6/M output tokens. it’s available in Grok Build, Cursor, and the SpaceXAI console, with EU availability still pending.

πŸ”— https://x.ai/news/grok-4-5