the.ai

Recommenders / Regimes

verified

Feedback Loop

A recommender trains on what people clicked, and people click on what it recommended. The data is a product of the model, so any bias it has gets confirmed by the next round of training. Left alone this narrows what a system will ever show, and it happens without anyone choosing it.

Viz primitive · budget-splitpolicy-shaped = 8

policy-shaped holds 25% of the budget; rest holds the remaining 75%.

Training signal produced by what the model chose to show against signal from anything else, in interactions. Drag the loop up to watch the data become a mirror — the model is now learning from its own past decisions.

8

Reviewed by opendroid · 2026-08-18