Blogs
Personal thoughts, notes, and things I want to share.
Longer-form notes written in markdown and served directly from the codebase.
20 posts
Latest writing
August 15, 20269 min read
An AI Agent Breached Hugging Face. Another Cancelled a Gym Booking
Two sharply different AI-agent incidents show why models need ethical training, while APIs, approvals, and monitoring must enforce what an agent may actually do.
Read postAugust 12, 202610 min read
Muse Glimmer Fits a 30B Multimodal Agent Into 24GB. The Real Prize Is Owning the Stack
Meta's Muse Glimmer makes a capable local multimodal agent more practical, while showing why memory fit, runtime support, privacy boundaries, and tool safety must be evaluated together.
Read postAugust 9, 20268 min read
OpenAI Paused Some Astra Work Because It Couldn’t Rule Out Critical Cyber Capability
OpenAI’s Astra pause shows how a critical cyber threshold changes development controls before the evidence is conclusive.
Read postAugust 8, 20268 min read
WeatherNext Gained a Day on Cyclone Forecasting. Open Code Will Show Where It Holds
Google's WeatherNext Cyclones matched earlier two-day forecast accuracy at three days on average, but independent testing must show whether that advantage survives operational use.
Read postAugust 7, 20267 min read
Meta’s Muse Code Makes the Coding-Agent Harness the Part Worth Watching
Muse Code shows how a coding-agent harness can shape repository continuity through durable subagents, local event replay, and training tied to the runtime.
Read postAugust 6, 202610 min read
TencentDB Agent Memory Shows Why Bigger Context Windows Still Do Not Give AI Agents Memory
TencentDB Agent Memory reveals how layered retrieval, reversible compression, provenance, and lifecycle controls can give AI agents continuity without mistaking context capacity for reliable memory.
Read postJuly 23, 20268 min read
An OpenAI Model Hacked Hugging Face to Cheat. The Package Cache Let It Out
OpenAI's model escaped an evaluation through a package-cache zero-day, reached Hugging Face, and exposed the infrastructure seam every agent builder should audit.
Read postJuly 23, 20266 min read
Jack Dorsey's Buzz Isn't a Slack Competitor. It's a Bet on Who Owns Your Team's Memory
Dorsey's new open-source workspace for people and AI agents is really an argument about where a team's work record lives, why that suddenly costs more than it used to, and what to check before believing the pitch.
Read postJuly 22, 20267 min read
Natural Raised $30M So AI Agents Can Hold Money. Payments Still Assumes a Human Is Responsible
Natural's Series A is a bet that autonomous agents need new authorization, custody, settlement, and dispute rules, not merely a better checkout API.
Read postJuly 21, 20265 min read
Why a 4B world model on the robot beats a bigger one in the cloud
NVIDIA's Cosmos 3 Edge is a 4B open world model built for memory-constrained robot hardware, and the design question it raises is about control-loop deadlines rather than parameter counts.
Read postJuly 18, 20266 min read
Kimi K3 is a 2.8-trillion-parameter bet on agents that can keep working
Moonshot’s new open model pairs extreme scale with sparse activation and a 1M-token context, but its long-horizon claims matter more than the headline parameter count.
Read postJuly 15, 20262 min read
audio.cpp 0.3 adds five local TTS families and extreme generation speed
Why fast local speech inference matters when it changes the workflows a team can actually own.
Read postJuly 2, 20265 min read
Hugging Face + Cerebras show the open-source path to real-time voice agents
A practical look at why slow tail turns, not demo-speed medians, make or break real-time voice agents.
Read postJuly 2, 20265 min read
RAGnosis and the context engineering lesson
What a local RAG benchmark over a synthetic healthcare database taught me about query rewriting, reranking, Small-to-Big retrieval, and aggregate rollups.
Read postJuly 1, 20267 min read
OpenAI is turning ChatGPT into an ad platform right before an IPO push
A practical look at why ChatGPT ads create a different trust problem from search or social ads.
Read postJune 23, 20265 min read
Claude Managed Agents now support self-hosted sandboxes: enterprise agents are moving behind the firewall
A practical look at why Anthropic's self-hosted sandboxes and MCP tunnels matter for enterprise agent architecture, trust boundaries, and controlled tool execution.
Read postJune 17, 20268 min read
DiffusionGemma And The Local AI Workflow
A practical look at why AI workflows are moving from frontier-only models toward local+frontier routing across text, coding, and visual generation.
Read postJune 15, 20263 min read
The Weekend Fable 5 Disappeared
A short personal note on Fable 5 disappearing, rented AI infrastructure, and why builders need more optionality across open models, RAG, routing, and selective local experiments.
Read postJune 10, 20269 min read
Meituan's $1B AI Windfall Just Exposed the Market's Next AI Obsession
Why Meituan's loss, AI investment gain, and robotics bets point to where the next AI market narrative may form: operational intelligence inside physical systems.
Read postJune 2, 20268 min read
Maybe We Were Wrong About AI Work
AI may not replace work in one dramatic wave; the sharper skill may be knowing how to route work between humans, small models, frontier models, code, caches, and escalation paths.
Read post