Notes on web data.
How AI agents read, extract and monitor the web — written for builders.
Firecrawl vs Jina Reader vs Fastcrawl: an honest comparison for 2026
Pricing math, latency, and feature surface — what you actually get from each scraping API, and when each one is the right pick.
What 'LLM-ready markdown' actually means (and how to get it)
Boilerplate stripping, JS rendering, token cost, and structure — the four things that decide whether your RAG pipeline eats good data or garbage.
Monitoring websites with AI agents: change detection that doesn't wake you at 3am
Scheduled re-scrapes, hashing, diffing and webhooks — a practical pattern for tracking competitors, docs, prices and changelogs.