Cover image

The Query Ran. The Answer Was Still Wrong: What EXYGEN Changes About Knowledge-Graph Interfaces

EXYGEN shows why fine-tuning-free knowledge-graph querying still needs explicit schema context, task examples, execution-based validation, and scalable metadata generation.

September 24, 2026 · 7 min · Zelina
Cover image

Trust Signals Don’t Check Themselves: Put Verification Before the Install

A pre-registered audit of AI coding assistants shows why software provenance controls need to live in the execution workflow, not in assumptions about model capability.

September 24, 2026 · 8 min · Zelina
Cover image

Zero Is Not Nothing: When an Inactive Token Still Changes the Model

OAttention shows how to make exact zero-vector tokens structurally inert—and why doing so at the attention layer is not enough.

September 24, 2026 · 6 min · Zelina
Cover image

A Green Check Is Not a Physical Verdict: What SimVerity Changes About Agent Deployment

SimVerity shows why deployment clearance must specify which property transferred, at which observation boundary, and with what physical evidence.

September 23, 2026 · 7 min · Zelina
Cover image

Before You Spend 10 Trillion Tokens: Separate Width From Horizon

A two-step learning-rate transfer method shows how MoE training teams can reduce full-scale tuning compute by separating model-width transfer from token-horizon extrapolation.

September 23, 2026 · 7 min · Zelina
Cover image

Give Sound a Sense of Space: Depth as a Constraint in Audio-Visual Segmentation

DGCM-AVS shows how estimated depth can constrain audio-visual localization when semantics alone leave multiple plausible sounding objects.

September 23, 2026 · 8 min · Zelina
Cover image

Parallel by Default: Manufacturing Agents Need Routing, Not One Reasoning Mode

Design-to-Plan shows why manufacturing agent systems may benefit from parallel execution for routine work while reserving deeper sequential reasoning for cases that demand validation.

September 23, 2026 · 7 min · Zelina
Cover image

Same Chain, Different Signal: Why Solana Rug-Pull Models Do Not Travel Cleanly

A large Solana study shows that five minutes of trading data can flag later rug-pull risk, but venue changes can erase much of that signal.

September 23, 2026 · 6 min · Zelina
Cover image

The Safety Leaderboard Has Conditions: Choose Moderators by Harm, Context, and Cost

A 53-model benchmark shows why content-moderation procurement should depend on harm type, observation point, statistical uncertainty, and deployment cost rather than one global ranking.

September 23, 2026 · 8 min · Zelina
Cover image

When the Proxy Wins: Attention Sensitivity Can Hit the Ceiling and Still Miss ICL

A controlled fine-tuning study shows why an internal model diagnostic can become a dangerous training target once optimization learns how to satisfy the metric without preserving the intended capability.

September 23, 2026 · 7 min · Zelina