Cover image

A Multilingual Research Assistant Is Still an Infrastructure Project

ReSearch_SSH shows how specialised research platforms can combine domain adaptation, behavioural retrieval, knowledge graphs, expert evaluation, and legal controls without discarding their existing infrastructure.

August 4, 2026 · 9 min · Zelina
Cover image

Attention Is a Connection Walk, Not Automatically a Laplacian

A precise operator view shows why attention maps capture token routing but omit the feature transformations that determine what the layer actually computes.

August 4, 2026 · 7 min · Zelina
Cover image

The Probe Saw the Prompt Before It Saw the Fake

A hidden-state monitor can detect alignment-faking signals, but only after controls separate strategic compliance from prompt identity, query leakage, and configuration artifacts.

August 4, 2026 · 8 min · Zelina
Cover image

Let the Model Design the Poster—Not the Evidence

PosterHarness shows how to make scientific poster generation auditable by separating visual composition from evidence-bearing figures.

August 3, 2026 · 8 min · Zelina
Cover image

The Catalog Grew. The Agent Needed a Call Stack.

A hierarchical agent architecture sharply reduces tool-schema exposure at scale, but only when taxonomy, validation, and latency controls are engineered with equal care.

August 3, 2026 · 9 min · Zelina
Cover image

The Memory Score Changed Before the Memory Did

MemTools shows why agent-memory evaluation must separate component quality from interface compatibility, execution timing, and representation coordination.

August 3, 2026 · 9 min · Zelina
Cover image

Stored Is Not Reachable: Why Continual Fact Writing Breaks Under Later Updates

Controlled experiments show that a fact can remain statistically present in an LLM’s weights while becoming inaccessible, unusable, and vulnerable to later updates.

August 2, 2026 · 10 min · Zelina
Cover image

The Model Saw Every Scene. The System Had to Remember the Story.

StoryTeller shows why coherent long-form narration depends on verified external state, not only a stronger vision-language model or a longer prompt.

August 2, 2026 · 8 min · Zelina
Cover image

The Prompt Knew the Odds. CRISTAL Put Them in Code

CRISTAL shows why analyst systems may need LLMs for interpretation but explicit probabilistic code for evidence weighting, updating, and final decisions.

August 2, 2026 · 8 min · Zelina
Cover image

Fair on Clean Data, Fragile After Fake Profiles

A recommender can pass a clean-data fairness review and still develop larger subgroup gaps after coordinated fake profiles enter its retraining data.

August 1, 2026 · 8 min · Zelina