Cover image

MoA vs. Moat: Agentic LLMs for Drug Competitor Mapping Cut Diligence Time 20×

A mechanism-first look at how scaffolded web agents and conservative validation turn messy biotech diligence into faster, more measurable competitor mapping.

August 25, 2025 · 17 min · Zelina
Cover image

Preference Chains of Command: Making LLM Agents Pick Like People

A mechanism-first reading of Preference Chain, a Graph RAG method that uses small behavioural samples to make LLM mobility agents less generic and more locally plausible.

August 25, 2025 · 21 min · Zelina
Cover image

Put It on the GLARE: How Agentic Reasoning Makes Legal AI Actually Think

GLARE shows that legal AI improves when it broadens candidate charges, learns from precedent reasoning paths, and searches only for missing legal premises.

August 25, 2025 · 17 min · Zelina
Cover image

ReAct Without the Chaos: AgentScope 1.0 Turns Tools into Strategy

AgentScope 1.0 shows how tool-using agents become more operationally credible when ReAct is wrapped in disciplined abstractions, tracing, evaluation, and runtime isolation.

August 25, 2025 · 17 min · Zelina
Cover image

Spin Doctors: Why RL Fine‑Tuning Mostly Rotates, Not Reinvents

A spectral reading of SFT and RL fine-tuning shows why RL often restores lost generalization rather than manufacturing new capability from scratch.

August 25, 2025 · 14 min · Zelina
Cover image

Charting a Better Bedside: When Agentic RL Teaches RAG to Diagnose

Deep-DxSearch shows that diagnostic RAG becomes more useful when retrieval behaviour is trained as a policy, not scripted as a prompt.

August 24, 2025 · 18 min · Zelina
Cover image

Enemy at the Gates, Friends at the Table: Why Competition Makes LLM Agents More Cooperative

A mechanism-first reading of why external rivalry can make LLM agent teams cooperate more internally, and why that lesson is useful but not yet operational proof.

August 24, 2025 · 19 min · Zelina
Cover image

From Tokens to Teaspoons: What a Prompt Really Costs

Google’s Gemini serving paper is less interesting as a tiny per-prompt footprint claim than as a practical accounting template for measuring AI inference.

August 24, 2025 · 17 min · Zelina
Cover image

Peer Review, But Make It Multi‑Agent: Inside aiXiv’s Bid to Publish AI Scientists

aiXiv shows that the hard part of AI-generated science is not generation, but building the review, revision, security, and governance machinery around it.

August 24, 2025 · 17 min · Zelina
Cover image

Stackelbergs & Stakeholders: Turning Bits into Boardroom Moves

BusiAgent shows how multi-agent LLMs can turn broad business requests into governed workflows, but its real value is orchestration discipline, not artificial executive genius.

August 24, 2025 · 18 min · Zelina