2026 · 04
Long-horizon reasoning without forgetting
A memory schedule that keeps week-scale context coherent while holding compute flat.

Research
We publish what we learn — including the parts that did not work.
Papers & notes
2026 · 04
A memory schedule that keeps week-scale context coherent while holding compute flat.
2026 · 02
Training signals that make models prefer verifiable claims over fluent ones.
2025 · 11
Why self-critique improves precision but collapses recall, and how to reopen it.
2025 · 07
Team-authored rubrics beat public benchmarks for predicting real-world usefulness.
Principles
They shape what we publish and what we refuse to ship.
If a result cannot be explained, it is not a result.
Any claim the system makes must be challengeable by a person.
Capability without a use we understand stays in the lab.