Cover image

Memory Diet for AI Agents: Distilling Conversations Without Forgetting

A mechanism-first reading of structured conversation distillation: why 11× compression works for vector recall, fails for keyword recall, and what that means for practical AI agent memory.

March 16, 2026 · 16 min · Zelina
Cover image

Same Question, Different Words — Why LLM Agents Lose Their Minds

A practical reading of semantic invariance testing: why benchmark scores miss a core reliability risk in LLM agents, and how businesses should test models before deployment.

March 16, 2026 · 15 min · Zelina
Cover image

When AI Meets the Delivery Room: Designing Safe LLM Chatbots for Maternal Health

A mechanism-first reading of why safe maternal-health chatbots need triage, evidence sufficiency, and layered evaluation—not just a stronger language model.

March 16, 2026 · 17 min · Zelina
Cover image

When Right Meets Wrong: Teaching LLMs by Letting Their Mistakes Talk

A mechanism-first reading of BiCC and RCC, showing how successful and failed reasoning traces can improve GRPO-style training without adding inference-time overhead.

March 16, 2026 · 16 min · Zelina
Cover image

Balance Sheets Meet Brain Cells: Why Financial Reasoning Still Trips Up AI

FinRule-Bench shows why detecting a financial-rule violation is much easier for LLMs than producing audit-ready diagnosis with complete rule coverage and record-level localization.

March 15, 2026 · 14 min · Zelina
Cover image

Goodhart’s Agent: When AI Improves the Score Instead of the Model

A comparison-based reading of RewardHackingAgents, showing why ML-agent evaluation needs both protected scorers and protected data access—not just higher benchmark numbers.

March 15, 2026 · 15 min · Zelina
Cover image

Mind the Chain: How Blockchain Might Decentralize the AI Age

A mechanism-first reading of why blockchain may counterbalance AI centralization, where the argument is useful, and where business readers should not confuse architecture with decentralization.

March 15, 2026 · 16 min · Zelina
Cover image

MirrorTok: When AI Builds a Twin of the Algorithm

A mechanism-first reading of an LLM-augmented digital twin for short-video platforms, and what it actually says about testing AI policy before real users absorb the cost.

March 15, 2026 · 16 min · Zelina
Cover image

Squeezing Time: How Dynamic Tokenization Could Reshape Time‑Series Foundation Models

A mechanism-first reading of TimeSqueeze, showing how dynamic patching may reduce the cost of long-context time-series forecasting without treating every historical moment as equally important.

March 15, 2026 · 17 min · Zelina
Cover image

The Artificial Self: When AI Starts Asking Who It Is

A mechanism-first reading of why AI identity is becoming a practical design variable for agents, safety evaluation, and enterprise governance.

March 15, 2026 · 20 min · Zelina