Cover image

Make the Structure Survive: Forma Turns Synthetic Clinical Cases Into Auditable Outputs

Forma shows that synthetic clinical cases become more controllable when the relationships that must survive generation are specified explicitly and audited after generation.

September 25, 2026 · 7 min · Zelina
Cover image

More Agents, More Rules: HiMA-MDD Treats Multi-Agent AI as a Governance Problem

HiMA-MDD shows that evidence permissions, decision ownership, bounded revision, and provenance can matter more than simply adding specialist agents.

September 25, 2026 · 6 min · Zelina
Cover image

Query Filters Catch the Attack That Advertises Itself

CamoDocs shows why RAG poisoning defenses must look beyond query overlap and compact embedding clusters—and why aggressive evidence removal can undermine the retrieval system they protect.

September 25, 2026 · 7 min · Zelina
Cover image

Teach the Restoration, Skip the Runtime Step

ReCAST shows how explicit de-obfuscation supervision can improve SMS risk detection without adding a restoration model to every production request.

September 25, 2026 · 7 min · Zelina
Cover image

The Linker Can’t Rank What It Never Sees

MACE shows that event-linking accuracy can improve by fixing the evidence used to build candidate sets before replacing or retraining the downstream linker.

September 25, 2026 · 7 min · Zelina
Cover image

The Model Felt the Tampering. It Couldn’t Name the Cause

A benchmark of output tampering shows why internal anomaly signals, visible model failure, and accurate self-diagnosis should be treated as separate production metrics.

September 25, 2026 · 7 min · Zelina
Cover image

Fast Without False Precision: Foundation Models for Partial Causal Identification

A causal foundation model can make repeated partial-identification queries fast without pretending observational data determine one causal answer.

September 24, 2026 · 6 min · Zelina
Cover image

More Data, Better Learner: What Children Reveal About Learning Efficiency

A child–language-model comparison suggests that data efficiency depends not only on how much input a learner receives, but on whether later experience becomes progressively more valuable.

September 24, 2026 · 7 min · Zelina
Cover image

New Signer, Old Sentence: What Sign-Language Translation Benchmarks Are Actually Testing

A signer-independent test can expose hidden robustness losses in sign-language translation, but it can still reward memorisation when target sentences repeat across splits.

September 24, 2026 · 7 min · Zelina
Cover image

The Multilingual Tax Is Logarithmic: What Actually Breaks as Language Coverage Grows

A theoretical result separates the modest representation cost of adding languages from the training-budget and evaluation choices that can make multilingual quality deteriorate in practice.

September 24, 2026 · 8 min · Zelina