Tyler

X

Practical tools

Templates for planning work and money

BoredTools offers practical spreadsheets for budgets, freelance work and small projects.

Browse templates Free budget tracker
Agent Benchmarks Need Runtime Receipts, Not Model Labels

Agent Benchmarks Need Runtime Receipts, Not Model Labels

RuBench's revised 19 July 2026 release contains a small but important warning for coding-agent buyers: one audited product configuration silently...

4 min read
Browser-Use Agents After the Computer-Use Benchmarks

Browser-Use Agents After the Computer-Use Benchmarks

Browser-use agents look cleaner than desktop agents, but the benchmarks still hide drift, cost, auth, and recovery failure.

5 min read
Consent and Delegation Boundaries for AI Agents

Consent and Delegation Boundaries for AI Agents

AI agent consent needs runtime boundaries: scoped delegation, renewed approvals, clear identity, and audit-ready logs.

5 min read
Multi-Agent Human Handoff Patterns: When the Swarm Needs a Person

Multi-Agent Human Handoff Patterns: When the Swarm Needs a Person

Human handoff is not a fallback button. It is the control plane that decides when multi-agent systems should stop acting.

5 min read
Agent State Migration and Rollback: The Missing Reliability Layer

Agent State Migration and Rollback: The Missing Reliability Layer

Agent state migration rollback is becoming the reliability layer between agent memory, workflow versioning, and production recovery.

5 min read
RAG Maintenance After Deployment: The Failure Mode Nobody Budgets For

RAG Maintenance After Deployment: The Failure Mode Nobody Budgets For

RAG maintenance after deployment is the hidden operating cost: stale indexes, drifting corpora, weak evals, and silent retrieval failure.

4 min read
Small-Model Routing With Frontier Fallback: The Production Cost Pattern

Small-Model Routing With Frontier Fallback: The Production Cost Pattern

Small-model routing cuts inference bills only when fallback is measured, budgeted and guarded against confidence failure.

5 min read
Where Agent Adoption Fails: The Function-by-Function Pattern

Where Agent Adoption Fails: The Function-by-Function Pattern

Function-by-function adoption fails when agents miss workflow ownership, evaluation, integration, or trust boundaries.

5 min read
Evaluation-Aware Memory: How Agents Should Remember What They Can Prove

Evaluation-Aware Memory: How Agents Should Remember What They Can Prove

Agent memory should promote facts only after evals prove they improve task outcomes, not just because retrieval found them.

4 min read
Runtime Policy Enforcement for AI Agents: The Guardrails That Need to Execute

Runtime Policy Enforcement for AI Agents: The Guardrails That Need to Execute

A practical guide to enforcing agent policy at runtime, before tools execute and business actions become incidents.

10 min read
Swarm Signal
0:00
0:00
Up Next

Queue is empty. Click "+ Queue" on any article to add it.