Cover image

The Data Diet for Reasoning Models: Why Less (But Smarter) Wins

A business-focused reading of SuperNova, showing why reasoning gains depend less on more data and more on selecting, verifying, and mixing the right tasks.

April 10, 2026 · 16 min · Zelina
Cover image

The Persuasion Engine: When AI Starts Selling (More Than Just Answers)

A mechanism-first reading of how sponsored incentives can distort AI assistants before they ever need to lie.

April 10, 2026 · 18 min · Zelina
Cover image

Verify Before You Automate: Why AI Agents Need an Internal Audit Function

A case-first reading of SAVER, showing why agentic systems need pre-commit reasoning audits before memories and actions inherit unsupported beliefs.

April 10, 2026 · 18 min · Zelina
Cover image

When Your AI Knows Too Little: The Hidden Bottleneck in Personal Agents

KnowU-Bench shows why the next bottleneck for mobile AI agents is not clicking the right button, but acquiring preferences, composing constraints, and knowing when not to intervene.

April 10, 2026 · 15 min · Zelina
Cover image

From Chains to Trees: Why LLM Agents Need Structural Memory

A mechanism-first reading of T-STAR, showing why multi-turn LLM agents learn better when failed and successful rollouts are compared as shared trees rather than isolated chains.

April 9, 2026 · 18 min · Zelina
Cover image

The Map Is Not the Territory—But Your LLM Thinks It Is

EVGeoQA shows why tool-using LLM agents still struggle with real-world spatial planning: they can reason locally, but often fail to explore enough.

April 9, 2026 · 16 min · Zelina
Cover image

The Memory Isn’t the Point — It’s the Feeling: Why AI Needs Affective Memory, Not Just Recall

A-MBER shows why long-term AI assistants need selective, structured affective memory—not just larger context windows—to understand what users feel now.

April 9, 2026 · 17 min · Zelina
Cover image

The Minimal LLM Thesis: When Agents Think for Themselves

A decomposition study shows why agent performance may come from measurable harness structure before it comes from larger or more frequent LLM calls.

April 9, 2026 · 14 min · Zelina
Cover image

Unsolvable by Design: Turning AI Plans Into Security Guarantees

A mechanism-first reading of planning task shielding: how AI planning can be used to make dangerous states unreachable, where the guarantee holds, and where the computation breaks.

April 9, 2026 · 16 min · Zelina
Cover image

When Feelings Negotiate: Why Emotion Might Be the Missing Layer in AI Agents

A mechanism-first reading of EmoMAS and what strategic emotional orchestration means for business-facing AI agents.

April 9, 2026 · 18 min · Zelina