▣ App Dev/2026-04-23Advanced
Semantic Caching for LLM Responses in Antigravity — Estimating Your Own Savings Before You Build It
Building a semantic LLM response cache with Antigravity, pgvector, and Gemini. Covers migrating off the retired text-embedding-004, normalizing truncated vectors, and a formula that turns hit rate and unit-cost ratio into an actual savings estimate.