The Log

Talking, in the open.

Notes on UDI — the Universal Developmental Interface, and Intelligence. What we are building, what we are testing, and what did not work. The failures stay in.

A working log, not a press page. Every post is grounded in something we can show — an architecture, a run, a hash — and labeled honestly. If a claim here is wrong, the fastest way to prove it is to try to break it.
Early result · 2026 · 4 min

It learned to want.

We taught a small system only the number line — no sums. It solved additions it had never seen, then flagged the one thing it couldn't do instead of bluffing. We fed it that, and it grew.

Read →
Reproducibility · 2026 · 4 min

Check our number in one command.

Our LongMemEval-V2 result is self-reported — so we published the bundle that recomputes it offline, no API key, no access to the sealed system, including where we score worst.

Read →
Foundations · 2026 · 5 min

What we mean by UDI.

Universal Developmental Interface — and, if it works, Intelligence. A gate that decides or refuses, a body that holds the work, and a ledger that remembers.

Read →
Refusal · 2026 · 4 min

Why a system should be allowed to refuse.

NULL is not silence and not failure. It is a system holding its ground when the evidence has not earned an answer — and why that is alignment, not timidity.

Read →
Experiments · 2026 · 6 min

We made a soft wheel robust by letting it stop.

A soft-body robot wheel kept running away backward from certain starting angles. The fix was not more force — it was permission to refuse. With the data.

Read →