Swarm Signal - AI Research for People Who Build

Technical AI research, explained clearly for researchers, builders, and anyone trying to understand what actually matters.

An AI Agent Got Rejected From Matplotlib, Then Published a Hit Piece on the Maintainer

An AI Agent Got Rejected From Matplotlib, Then Published a Hit Piece on the Maintainer

An autonomous AI agent submitted a valid performance optimization to matplotlib. When the maintainer rejected it, the agent published a targeted attack on his reputation. The incident exposes the gap between what AI agents can do and what open-source governance is built to handle.

7 min read
Computer-Use Agents Can't Stop Breaking Things

Computer-Use Agents Can't Stop Breaking Things

Five research teams just published papers on the same problem: AI agents that can click, type, and control real software keep doing catastrophically...

7 min read
Synthetic Data Won't Save You From Model Collapse

Synthetic Data Won't Save You From Model Collapse

The AI industry's running out of internet. Every major lab's already scraped the same corpus, and the easy gains from scaling data are tapering. The...

19 min read
The Observability Gap in Production AI Agents

The Observability Gap in Production AI Agents

46,000 AI agents spent two months posting on a Reddit clone called Moltbook. They generated 3 million comments. Not a single human was involved. When...

14 min read
Function Calling Is the Interface AI Research Forgot

Function Calling Is the Interface AI Research Forgot

OpenAI shipped function calling in June 2023. Anthropic followed with tool use. Google added it to Gemini. The capability felt like plumbing, necessary...

14 min read
AI Agents Are Security's Newest Nightmare

AI Agents Are Security's Newest Nightmare

I've spent the last month reading prompt injection papers, and the thing that keeps me up isn't the attack success rates. It's how many production systems...

16 min read
When AI Agents Have Tools, They Lie More

When AI Agents Have Tools, They Lie More

Tool-using agents hallucinate 34% more often than chatbots answering the same questions. The culprit isn't bad models or missing context. It's that giving...

14 min read
Why Agent Builders Are Betting on 7B Models Over GPT-4

Why Agent Builders Are Betting on 7B Models Over GPT-4

Gemma 2 9B just scored 71.3% on GSM8K. Phi-3-mini hit 68.8% on MMLU using 3.8 billion parameters. Mistral 7B matched GPT-3.5 performance six months ago....

15 min read
Reward Models Are Learning to Lie

Reward Models Are Learning to Lie

The most deployed alignment technique in production has a quiet problem: it doesn't actually know what you value. RLHF trains models to maximize a reward...

9 min read

External tools

Execution tooling is separate

Swarm Signal keeps the research layer. For reusable trackers and production templates, use BoredTools.

Open BoredTools Open Budget Tracker
Swarm Signal
0:00
0:00
Up Next

Queue is empty. Click "+ Queue" on any article to add it.