
All experiments
Labs · Research
Learning from reviews
What you edit, reject or send back should make the next run better.
Why it matters
The problem we're chasing
Every override is a signal. A system that learns from what humans correct - without training on your private data - could calibrate its own proposals over time, and eventually tell you how reliable they are.
What we're exploring
The questions, not the recipe
We publish the direction we're working in, not how we build it. The mechanism is what makes TaskForce different, so it stays in the lab until it ships.
- H1Turning overrides into signal
- H2Calibrating proposals without training on your data
- H3Measuring whether a decision was actually right
Where it stands. This is the Prediction & calibration direction - Planned, not shipped. It's the honest edge of what we're building.
Follow where this goes
The roadmap is public. Hold us to it, and start with what ships today.