Practical tools
Templates for planning work and money
BoredTools offers practical spreadsheets for budgets, freelance work and small projects.
AI Coding Agents: What Actually Works in Production
GitHub reports that 46% of all new code is now AI-generated. Ninety-two percent of US developers use AI coding tools daily. Claude Code hit $2.5 billion...
Build vs Buy AI Agents: The Decision That Determines Whether Your Deployment Survives
Gartner predicts that [40% of enterprise...
Small Agent Models Need Tool Floors, Not Parameter Claims
Small language models are getting a serious agent story, but the useful question is no longer whether a 1B, 3B, or 8B model can sound capable. The useful...
Model Selection Guide: How to Pick the Right AI Model for Your Use Case
A March 2026 survey of the [Artificial Analysis leaderboard](https://artificialanalysis.ai/) counts 429 tracked models, over 200 of them open-weight....
Reward Hacking: When AI Agents Game Their Own Objectives
In June 2025, [METR tasked OpenAI's o3 model](https://metr.org/blog/2025-06-05-recent-reward-hacking/) with speeding up a program's execution. Instead of...
Agent Messages Need State, Not Chat
Multi-agent systems do not only fail because the agents are weak. They also fail because every agent is allowed to narrate too much. A June 2026 paper...
Seven Protocols, 1% Adoption: The Agent Economy's Infrastructure-Reality Gap
Visa, Mastercard, PayPal, Stripe, Coinbase, Google, and Shopify all shipped agent payment protocols in the last sixteen months. Seven competing standards...
Your Agent Doesn't Need Human Memory. It Needs Something Weirder.
The AI industry keeps describing agent memory like it's a brain. "Short-term memory," "long-term memory," "episodic recall." The metaphors are intuitive....
AI Interpretability Tools in 2026: What the Research Actually Shows
▶️ LISTEN TO THIS ARTICLE Your browser does not support the audio element. AI Interpretability Tools in 2026: What the Research Actually Shows Interpretability is one part of a broader debugging stack. For teams building AI agents, a practical question is which tools help debug a failure, inspect behavior, or monitor
Test-Time Compute in 2026: The Complete Practitioner's Guide
The new frontier in AI performance isn't bigger models. It's smarter inference. Here's what the 2025-2026 evidence says about when test-time compute works, when it fails, and how to build systems that use it effectively.