Cover image

Parallel Worlds of Moderation: How LLM Simulations Are Stress-Testing Online Civility

COSMOS turns online moderation into a counterfactual simulation problem, showing why personalised interventions may reduce toxicity without the collateral damage of blunt bans.

November 12, 2025 · 16 min · Zelina
Cover image

Patch, Don’t Preach: The Coming Era of Modular AI Safety

A mechanism-first look at safety policy patching, a lightweight way to update LLM safety behaviour without redeploying full model weights.

November 12, 2025 · 18 min · Zelina
Cover image

Proof, Policy, and Probability: How DeepProofLog Rewrites the Rules of Reasoning

DeepProofLog reframes symbolic proof search as policy learning, showing how neurosymbolic AI can scale reasoning without throwing away proof-level interpretability.

November 12, 2025 · 18 min · Zelina
Cover image

The Gospel of Faithful AI: How FaithAct Rewrites Reasoning

FaithAct turns multimodal reasoning from fluent narration into evidence-checked planning, making hallucination less a personality flaw and more an engineering defect.

November 12, 2025 · 14 min · Zelina
Cover image

The Problem with Problems: Why LLMs Still Don’t Know What’s Interesting

A study of math-problem interestingness shows why AI systems need calibrated taste, not just stronger solving ability.

November 12, 2025 · 15 min · Zelina
Cover image

DeepPersona and the Rise of Synthetic Humanity

DeepPersona shows that synthetic users become useful not by becoming longer, but by becoming structured, controllable, and empirically testable.

November 11, 2025 · 18 min · Zelina
Cover image

Forget Me Not: How IterResearch Rebuilt Long-Horizon Thinking for AI Agents

IterResearch shows why long-horizon AI agents need disciplined workspace reconstruction, not merely longer context windows.

November 11, 2025 · 17 min · Zelina
Cover image

Parallel Worlds of Moderation: Simulating Online Civility with LLMs

A mechanism-first reading of COSMOS, an LLM-powered counterfactual simulator for testing moderation strategies before exposing real communities to policy experiments.

November 11, 2025 · 18 min · Zelina
Cover image

Touch Intelligence: How DigiData Trains Agents to Think with Their Fingers

A mechanism-first reading of DigiData, Meta’s dataset and benchmark for training mobile agents to complete real app tasks rather than merely imitate taps.

November 11, 2025 · 17 min · Zelina
Cover image

When Agents Think in Waves: Diffusion Models for Ad Hoc Teamwork

A mechanism-first reading of PADiff, showing why diffusion policies may help agents preserve multiple cooperation plans when working with unfamiliar teammates.

November 11, 2025 · 18 min · Zelina