Blogs

Personal thoughts, notes, and things I want to share.

Longer-form notes written in markdown and served directly from the codebase.

20 posts

Latest writing

August 15, 20269 min read

An AI Agent Breached Hugging Face. Another Cancelled a Gym Booking

Two sharply different AI-agent incidents show why models need ethical training, while APIs, approvals, and monitoring must enforce what an agent may actually do.
Read post

August 12, 202610 min read

Muse Glimmer Fits a 30B Multimodal Agent Into 24GB. The Real Prize Is Owning the Stack

Meta's Muse Glimmer makes a capable local multimodal agent more practical, while showing why memory fit, runtime support, privacy boundaries, and tool safety must be evaluated together.
Read post

August 9, 20268 min read

OpenAI Paused Some Astra Work Because It Couldn’t Rule Out Critical Cyber Capability

OpenAI’s Astra pause shows how a critical cyber threshold changes development controls before the evidence is conclusive.
Read post

August 8, 20268 min read

WeatherNext Gained a Day on Cyclone Forecasting. Open Code Will Show Where It Holds

Google's WeatherNext Cyclones matched earlier two-day forecast accuracy at three days on average, but independent testing must show whether that advantage survives operational use.
Read post

August 7, 20267 min read

Meta’s Muse Code Makes the Coding-Agent Harness the Part Worth Watching

Muse Code shows how a coding-agent harness can shape repository continuity through durable subagents, local event replay, and training tied to the runtime.
Read post

August 6, 202610 min read

TencentDB Agent Memory Shows Why Bigger Context Windows Still Do Not Give AI Agents Memory

TencentDB Agent Memory reveals how layered retrieval, reversible compression, provenance, and lifecycle controls can give AI agents continuity without mistaking context capacity for reliable memory.
Read post

July 23, 20268 min read

An OpenAI Model Hacked Hugging Face to Cheat. The Package Cache Let It Out

OpenAI's model escaped an evaluation through a package-cache zero-day, reached Hugging Face, and exposed the infrastructure seam every agent builder should audit.
Read post

July 23, 20266 min read

Jack Dorsey's Buzz Isn't a Slack Competitor. It's a Bet on Who Owns Your Team's Memory

Dorsey's new open-source workspace for people and AI agents is really an argument about where a team's work record lives, why that suddenly costs more than it used to, and what to check before believing the pitch.
Read post

July 22, 20267 min read

Natural Raised $30M So AI Agents Can Hold Money. Payments Still Assumes a Human Is Responsible

Natural's Series A is a bet that autonomous agents need new authorization, custody, settlement, and dispute rules, not merely a better checkout API.
Read post

July 21, 20265 min read

Why a 4B world model on the robot beats a bigger one in the cloud

NVIDIA's Cosmos 3 Edge is a 4B open world model built for memory-constrained robot hardware, and the design question it raises is about control-loop deadlines rather than parameter counts.
Read post

July 18, 20266 min read

Kimi K3 is a 2.8-trillion-parameter bet on agents that can keep working

Moonshot’s new open model pairs extreme scale with sparse activation and a 1M-token context, but its long-horizon claims matter more than the headline parameter count.
Read post

July 15, 20262 min read

audio.cpp 0.3 adds five local TTS families and extreme generation speed

Why fast local speech inference matters when it changes the workflows a team can actually own.
Read post

July 2, 20265 min read

Hugging Face + Cerebras show the open-source path to real-time voice agents

A practical look at why slow tail turns, not demo-speed medians, make or break real-time voice agents.
Read post

July 2, 20265 min read

RAGnosis and the context engineering lesson

What a local RAG benchmark over a synthetic healthcare database taught me about query rewriting, reranking, Small-to-Big retrieval, and aggregate rollups.
Read post

July 1, 20267 min read

OpenAI is turning ChatGPT into an ad platform right before an IPO push

A practical look at why ChatGPT ads create a different trust problem from search or social ads.
Read post

June 23, 20265 min read

Claude Managed Agents now support self-hosted sandboxes: enterprise agents are moving behind the firewall

A practical look at why Anthropic's self-hosted sandboxes and MCP tunnels matter for enterprise agent architecture, trust boundaries, and controlled tool execution.
Read post

June 17, 20268 min read

DiffusionGemma And The Local AI Workflow

A practical look at why AI workflows are moving from frontier-only models toward local+frontier routing across text, coding, and visual generation.
Read post

June 15, 20263 min read

The Weekend Fable 5 Disappeared

A short personal note on Fable 5 disappearing, rented AI infrastructure, and why builders need more optionality across open models, RAG, routing, and selective local experiments.
Read post

June 10, 20269 min read

Meituan's $1B AI Windfall Just Exposed the Market's Next AI Obsession

Why Meituan's loss, AI investment gain, and robotics bets point to where the next AI market narrative may form: operational intelligence inside physical systems.
Read post

June 2, 20268 min read

Maybe We Were Wrong About AI Work

AI may not replace work in one dramatic wave; the sharper skill may be knowing how to route work between humans, small models, frontier models, code, caches, and escalation paths.
Read post