If you are building a RAG or agent pipeline over video, the transcript is the payload. Titles and descriptions are thin; a 40-minute talk is thousands of tokens of dense, quotable prose. The good news is that YouTube already exposes that prose as captions. The bad news is that pulling captions for videos you do not own, at batch scale, is where most naive pipelines fall over. This post walks th...
Introduction
Every on-call engineer has had the same feeling: an alert fires, the symptoms look familiar, and you can't remember where you saw them before. Somebody fixed this months ago, but the fix lives in a closed ticket, a Slack thread, or one person's head.
I built RecallOps, an AI incident-response copilot, around that problem. It keeps a persistent recor...
Every tuning change needed a code review, a release, and a week of store rollout. So we stopped tuning, which is the worst possible outcome.
Originally published at guushu.com/notes. I keep the original updated, so this copy may lag.
Email Routing forwards a message and keeps nothing. If you want a copy you can search late...
My iPad lives on a stand most of the day. Videos, presentations, docs, whatever the work needs.
That is fine until I want a capture of whatever is on the screen. Taking the tablet off the stand for a few seconds sounds trivial, but it breaks the layout, moves the camera angle if I am recording, and usually means I lose the exact frame I meant to keep.
Apple already has a hardwar...
When a RAG system confidently states last quarter's price, last year's policy, or a fact that stopped being true in March, the reflex is to blame the model. Often the model did exactly its job: it grounded its answer faithfully in what retrieval handed it, and what retrieval handed it was stale. Your vector index is a cache of the world, and like any cache it goes wrong not by erroring but ...