Cover image

Wrong Code, Right Data Budget

TL;DR for operators If verified circuit-training data are scarce, exhaustive validation of every generated design may not be the best use of the next unit of data-engineering budget. In one experiment, a curated 22-circuit synthetic corpus reached 48.49% F1-Micro and 42.82% F1-Macro with a frozen circuit encoder, outperforming the original 22 verified circuits on both metrics and a 110-circuit raw generated corpus on F1-Micro. ...

September 4, 2026 · 7 min · Zelina
Cover image

Clean Less, Route Better: DataOrchestra Reframes Pretraining Data Curation

TL;DR for operators A pretraining-data pipeline receives millions of uneven records. Some are unusable, some contain removable noise, some need structural repair or added explanation, and some are already valuable enough that further processing may damage them. The operational problem is therefore not how to apply more cleaning, but how to decide which intervention—if any—each example needs. ...

August 9, 2026 · 8 min · Zelina
Cover image

The Data Diet for Reasoning Models: Why Less (But Smarter) Wins

A model-training team has a familiar bad habit: when the model fails, it asks for more. More examples. More domains. More synthetic prompts. More compute. More benchmarks to average over until the unpleasant details become small enough to ignore. This habit is understandable. It is also expensive. And, according to SuperNova, it may be the wrong first instinct. ...

April 10, 2026 · 16 min · Zelina
Cover image

When Maps Start Thinking: Teaching Agents to Plan in Time and Space

A map query is easy: get me from A to B. A service request is harder: leave after lunch, avoid tolls, find a charging station before the battery becomes theatrical, stop somewhere quiet for dinner, and make sure the restaurant is still open when we arrive. Every additional clause turns a lookup into a sequence of commitments. Locations must be resolved. Routes must be calculated. Opening hours, traffic, weather, prices, and travel times must remain mutually consistent. An incorrect essay can still sound intelligent. An incorrect itinerary can leave someone beside a closed charging station. ...

January 1, 2026 · 16 min · Zelina
Cover image

Eight Arms, One Mind: How OctoMed Turns Data Recipes into Medical Reasoning Power

Eight Arms, One Mind: How OctoMed Turns Data Recipes into Medical Reasoning Power Recipe sounds like a small word for an expensive problem. In medical AI, the usual boardroom story is simple: buy a bigger model, add more compute, sprinkle in reinforcement learning, and wait for clinical intelligence to appear. Very elegant. Also very convenient for anyone selling compute. ...

December 1, 2025 · 18 min · Zelina