Prompt Caching In Agents | EARENDIL
How prompt caching shapes the cost, latency, tools, and architecture of coding agents, and what Pi does to keep cache behavior visible.
How prompt caching shapes the cost, latency, tools, and architecture of coding agents, and what Pi does to keep cache behavior visible.
Personal site and blog of Tiago Araújo, software engineer working in AI/ML and full-stack.
written by Roni Kobrosly on 2026-07-23 | tags: engineering statistics human data
Open-weight models are becoming the foundation for the next AI ecosystem. The US should compete in it, not wall itself off.
Calibrate your enthusiasm
Approaches agents use to manage context: compaction, external retrieval, and learned experience.
The Whole Premise Of Checking For Human Writing Is Daft
Authorship lives in ideas, judgment, and responsibility, not in whether every word was typed by hand.
Uplifting a lobsters comment for easier reference.
A qualitative case study of how sociotechnical system design results in vulnerable populations
where delays are shown to be counterintuitive.
Changes proposed for ADB could kill an entire ecosystem of open-source power-user apps, mobile developer setups, and rootless privacy tools based on Shizuku.
Comparison and analysis of AI models across key performance metrics including quality, price, output speed, latency, context window & others.
Postgres LISTEN/NOTIFY Actually Scales | DBOS
How we optimized Postgres LISTEN/NOTIFY-backed data streams at scale, achieving 60K writes per second on a single Postgres server with millisecond-scale latency.