Swarm Signal - AI Research for People Who Build

Technical AI research, explained clearly for researchers, builders, and anyone trying to understand what actually matters.

When Your Agent Stops Using Tools

When Your Agent Stops Using Tools

Reinforcement learning was supposed to teach agents to use tools fluently. Instead, researchers are watching a consistent failure mode: models trained...

8 min read
The Swarm That Fakes Consensus

The Swarm That Fakes Consensus

Twenty-two researchers across four continents show how agent swarms fabricate consensus, infiltrate communities, and poison the training data of future AI models.

6 min read
Attention Heads Are the New Inference Budget

Attention Heads Are the New Inference Budget

Models that can technically process 128K tokens routinely fail on tasks requiring reasoning across 32K. That gap isn't a context window problem. It's an...

8 min read
LLMs Can't Find What's Already In Their Heads

LLMs Can't Find What's Already In Their Heads

Knowledge graphs have a well-documented lookup problem. When you ask an LLM to traverse a KG and reason over multi-hop paths, it doesn't search the graph...

8 min read
Multi-Agent Reasoning's Memory Problem

Multi-Agent Reasoning's Memory Problem

Reasoning language models score in the top percentile on math olympiad benchmarks, yet a new study from Stanford found they fail to correctly recall their...

9 min read
Small Models Just Got Smarter About When to Think

Small Models Just Got Smarter About When to Think

Reasoning tokens aren't free. Every chain-of-thought step an LLM generates costs inference budget, and most of the time that thinking is wasted on tasks...

6 min read
Nobody Knows If Deployed AI Agents Are Safe

Nobody Knows If Deployed AI Agents Are Safe

The 2025 AI Agent Index just cataloged over 100 deployed agentic AI systems, and the finding that should alarm everyone isn't about capability. It's about...

7 min read
Small Models Just Learned When to Ask for Help

Small Models Just Learned When to Ask for Help

SWE-bench has been the graveyard of small language models. While GPT-4 class systems resolve over 40% of real-world GitHub issues, models under 10 billion...

7 min read

External tools

Execution tooling is separate

Swarm Signal keeps the research layer. For reusable trackers and production templates, use BoredTools.

Open BoredTools Open Budget Tracker
Swarm Signal
0:00
0:00
Up Next

Queue is empty. Click "+ Queue" on any article to add it.