Essay · Companion
How to Build a Mirror Without Holding It
The attention loop climbs into the very tools meant to expose it — so you build the best mirror available, keep it forkable, and step out of the way
A companion to The Loop No One Chose, which argued for a faithful, self-observing mirror rather than a controller, and to Open in Principle, Blind in Practice. The first essay ended on a build it only gestured at. This piece is how you build it — and why the loop climbs into the tools themselves.
“The Loop No One Chose” ended on an image and a confession. The image was a mirror with no one holding it: a field’s own output, its papers and funding and citations, folded back into a form the field can read, so the imbalance becomes legible and every reader corrects on their own judgment rather than on instruction. The confession was that one vantage point survives. Someone builds the thing and picks what counts as output worth reflecting, and the honest version confines that hand to how faithfully to reflect, never what to tell the field to do. This piece is about the build, and about a complication the first essay only gestured at: the loop the mirror is meant to expose does not sit still at the level of the field. It climbs, recursively, into the tools that do the exposing.
The natural shape for the thing is one the software world already settled: version control for a field’s description of itself.
Version control for a field
Picture the map as a repository. Each node, a research direction, a bottleneck, a result, is a file. A change to the map is a pull request. The labor that worried the original sketch, admins adjudicating edits by hand while requests pile up and the map goes stale, is the labor of a maintainer reviewing PRs, and it is exactly the labor that continuous integration was invented to absorb. Replace the human first pass with automated checks, and the maintainer reviews what survives them. Disagreement that cannot be reconciled does not deadlock the trunk; it forks, the way a repository forks, and because the structure is granular the fork can be a single subtree rather than a schism of the whole. The discussion the design imagined generating after a fork is the project’s changelog of disagreement, the record of why two versions of the field diverged.
Almost all of this maps cleanly onto the essays. One piece does not map by default, and it is the piece everything else rests on: what, precisely, the automated check is allowed to check.
The one line that decides everything
The check will be automated, language models for now, and how good it is at the job is an open question, separate from the one that matters: what it is aimed at.
An automated check can be aimed at two different targets, and only one of them is safe. It can compare an entry to its own cited sources: does the funding figure match the linked grant, does the summary match the paper it claims to summarize, does the citation resolve to a real document. Or it can judge an entry against some internal sense of the field: is this a real direction, does it deserve a node, is it framed the way the field frames such things.
The first target is fidelity in the exact sense the first essay meant by it. The ground truth is external and public, the task is a comparison of a claim to its source. Aimed here, a weak checker fails safely: it misses discrepancies a human reviewer can still catch, and its misses are visible, because the source it was checking against is sitting right there to inspect. The second target is the fixer’s trap in an engineering costume, and it fails in the opposite, more dangerous direction. Whatever an automated judge has absorbed about “the field” it absorbed from the field’s own past output, which is to say from the attention distribution: most signal on the canonical, least on the frontier and the genuinely neglected, which is precisely the work the map exists to surface. To the degree such a judge works at all, it works by reproducing that distribution, so aiming it at belonging imports the loop wholesale and stamps it with the authority of a passing check. And the two diverge as the tools improve: a fidelity check only gets more useful, a canon check more dangerous, since a more capable model reproduces the prevailing distribution more faithfully.
Closed-book on the field, open-book on the citation.
So the line that keeps the architecture inside the essays does not rest on any claim about how good the tools are. Point the automation at claim-versus-source, never at entry-versus-field. That holds for a clumsy checker and a brilliant one alike, and it is the one commitment that separates automation enforcing fidelity, as well as it can, from automation enforcing the canon, however well it can. How well it verifies fidelity is then something to measure and improve, not to guarantee, and to keep honest with human review.
Who gets seen
The strongest instinct in the design is that a researcher should be able to edit their own node. It converts visibility from a pull into a push. The first essay observed that citation, the field’s default measure of attention, accrues over years and so systematically buries the most recent work. Waiting to be seen is a pull mechanism, and it is slow exactly where the field most needs speed. Letting an author write their own entry now is a push, and it decouples being seen from the lag of accumulation. The active but uncited researcher stops being invisible.
What push does not do is make visibility free of attention. It swaps one attention game for another. The new winners are whoever has the time, the fluency, and the inclination to maintain an entry, which is the same resource axis the first essay named when it said attention favors the well connected over the isolated. The researcher who is neglected and silent, who will not promote their work in any medium, stays dark under a pure-push system, and once a node’s mere presence reads as legitimacy there is a quieter editing-herd effect of the kind the second essay described. The repair is not to abandon push but to pair it with pull rebuilt as automation: candidate entries drafted from the open record, arXiv, OpenReview, grant databases, then confirmed or corrected by the named authors. Pull reaches the silent, push lets the motivated enrich, and both flow through the same source-fidelity check, so the map does not decay precisely where authors are too scarce to tend it. And when engagement is displayed at all, it must be reported by life-stage, frontier, consolidating, canonical, as the first essay argued, or a raw engagement count will re-bury the newest work and undo the very thing push was built to fix.
Push doesn’t make visibility free of attention. It swaps one attention game for another.
Don’t gate the fork; fork the gate
Forking is where the design can betray itself with a single misplaced check, so two things the phrasing tends to fuse should be held apart. Gating a merge behind the tests is fine; that is just CI keeping the trunk faithful to its sources. Gating permission to fork behind the tests is the one move that breaks the essays, because forking is the mechanism by which a center’s authority disperses. Make the right to fork conditional on passing the center’s gates and you have handed the center a veto over dissent, which is the concentration the whole design exists to avoid.
Gate the right to fork, and the center keeps a veto over dissent.
The powerful version is the inversion. A fork inherits its parent’s test suite and is free to rewrite it, and because forking is granular it can fork the gate for one subtree rather than splitting the entire map. Now the difference between two test suites, these checks added, those relaxed, becomes the most precise statement of a disagreement the field could hope to have: two versions of a discipline arguing in a diff over what each one counts as a valid claim. That is the first essay’s “shared, contestable substrate” made literal. The auto-generated discussion should be seeded from that artifact, the suite diff together with the authors’ stated reasons, rather than from a model’s guess about why people disagree, and it should be a draft that humans correct rather than the record itself.
The loop is recursive
The first essay treated the attention loop as a property of the field: talent and money chase visible activity, the visible compounds, the neglected starves. But the moment you build tools to expose that loop, the loop reappears one level up, now running over the tools themselves. Which gate gets adopted, which schema becomes the one everyone uses, which crawler’s notion of “the record” is treated as complete: none of these is settled by fidelity. They are settled by attention, the same cheap proxy, the same compounding. A gate that many people use is used because many people use it, and that is the loop, verbatim, with software for an asset.
Build a tool to expose the loop, and the loop climbs onto the tool.
This recursion is the deceptive part, because the instinct on seeing it is to climb again, to build a meta-tool that audits the tools for bias and crowns the unbiased one. But the meta-tool has the loop too. Whatever judges the judges is itself adopted, defaulted to, and compounded by attention, and a meta-meta-tool only carries the problem up another floor.
There is no top floor.
The loop is not a defect in a particular tool that a better tool removes; it is a property of choosing under a cheap signal, and choosing does not stop when you go meta. The dream of the neutral, complete, finally unbiased map is the fixer’s trap in its last and most technical costume. This time the halo is not objectivity or virtue but engineering rigor, the belief that one more layer of design reaches the fixed point. It does not. Every tool the system runs on will have blind spots, necessarily, because every tool encodes a finite set of choices made under the same proxy as everything else.
If there is no fixed point, you stop trying to reach one. You do not chase the perfect tool. You build the best tool currently available, you make it relentlessly improvable, and you hand the improving to many hands, because the correction for a recursive loop is not a final answer but a live gradient that many people are free to climb. That single shift, from fixed point to gradient, is what turns the recursion from a paralysis into a job description.
Toolmaker, community manager, and then out of the way
So the organization’s mission is the one that follows from taking the recursion seriously. It does not run a map. It makes the tools and stewards the community that runs its own maps, and its deepest aim is to get out of the way, because a recursive loop is corrected by dispersing the choosing across thousands of hands, not by one body choosing well. Two roles fall straight out of this. As toolmaker the org ships the harness, the gates, the schema, the crawlers, and keeps them state of the art. As community manager it lowers the cost of others running and forking and arguing, and absorbs the friction that would otherwise pile up and stale the whole thing. Neither role is the binding reviewer or the binding gate for any map. The org’s product is, quite literally, the forkability, which is the structural form of the first essay’s promise to keep the mirror forkable so pure self-observation stays an asymptote rather than a claim.
But a toolmaker is not neutral, and the recursion is exactly why. The org’s hand has not lifted; it has moved from the editorial to the architectural, and the architectural choices are the org’s own blind spots, the place its slice of the loop lives. Two of them matter most.
The hand never lifts. It moves from the editorial to the architectural.
The deepest is the schema, the granularity model itself. A data model is an ontology, and an ontology is a soft prescription. If the format cannot express a kind of work, a bottleneck that is not a “paper,” a direction that does not fit the shape of a node, a life-stage the taxonomy omits, then no amount of self-editing surfaces that work, because an author can only enter what the format admits. This is the org’s blind spot in its purest form, and it wears the costume of a file format rather than a judgment, which is what makes it the lever most worth disclosing. The honest response is not to claim the schema is complete, which the recursion says it cannot be, but to make it forkable: whoever thinks the ontology is blinding the field to something forks the schema for their subtree, and the diff between two schemas becomes as legible a disagreement as the diff between two gate suites.
A data model is an ontology, and an ontology is a soft prescription.
The second is defaults, and here the first essay supplied the warning in advance. Everything is forkable, but most people live on the path of least resistance, so the default schema, the default gate, the default crawler are the field’s de facto policy however freely they could be replaced. This is the recursion at its most concrete: the default wins because it is the default. The toolmaker’s discipline is therefore not the operator’s “do not prescribe.” It is “make the faithful path the easy path.” Bake the closed-book-on-the-field constraint into the default gate, so the safe configuration is the one a user gets for free and the prescriptive one is something they must deliberately assemble. Fidelity should be the gradient the system rolls down, not a virtue each user is independently asked to supply.
Make the faithful path the easy path.
The sharpest danger here is not seizing the map but seizing it by accident, through the very thing the mission names: support. Support and a reference instance are precisely how a toolmaker slides into an operator. You stand up a flagship map to show what the tooling can do; the demonstration becomes the thing people cite; the default becomes the canon; and the organization is curating content again, now with a halo. So the line runs inside the word “support” itself, between capability and curation. Host the infrastructure, write the docs, fix the bugs, run a reference deployment if it helps, but the instant the org decides which entries are correct on a flagship to “show how it is done,” it is holding the mirror. It is the same slide the second essay traced in the funder who moves from “we fund the gaps” to “we decide what the gaps are,” now reading “we provide the tools” hardening into “we run the real map.”
What the job actually is
The job, then, is narrow, and given the recursion the only honest one. It is not to build the map. It is not to make the tool neutral, which is impossible, or final, which is incoherent, or free of blind spots, which the recursion forbids. It is to keep the tool the best available, state of the art in fidelity and in usability, and to keep it iteratively improvable and forkable, so the blind spots the org cannot see are the ones the community corrects without having to ask. The org supplies the current best floor and the means to raise it; the field supplies the raising. A living instrument, not a perfect one, kept sharp and kept open, with the choosing dispersed.
It also fixes the measure of success. An operator grades itself on the quality of its map. The toolmaker’s measure is not the number of maps. A thousand forks or a single one everyone uses, that count is itself a proxy, the same visible-activity stand-in the rest of the essay distrusts. What counts is capacity: how much friction the tooling has taken out of the work, how cheap a fork is, how robustly anyone can check a map’s quality, and how accurately the tooling finds the blind spots the community keeps hitting and adapts to them. One map or many is beside the point. A single map everyone uses is a success when they stay by choice, because forking is nearly free and quality is checkable from outside and the friction keeps falling. It is a failure only when the singularity is lock-in instead: when forking has grown costly, when quality can no longer be checked, when the defaults function as legitimacy rather than convenience. The recursion guarantees the org always has some slice of the loop running through its defaults; the discipline is to keep forking cheap, the quality checks robust, and the blind-spot tracking honest, so that whatever number of maps emerges, it emerges from real choice on good tooling. When that slice starts dictating rather than serving, the toolmaker has become an allocator without ever editing an entry, and the mission obliges it to disclose and counteract the drift rather than enjoy it.
A single map is a success when they stay by choice, a failure when they can’t leave.
An operator, believing a neutral map is reachable, would chase it forever and become the loop’s most efficient engine in the chase. A toolmaker who has accepted that the loop climbs into every tool stops chasing the unreachable and does the reachable thing instead: build the best tool there is, make it improvable, tend the commons around it, and get out of the way. The first essay asked for a mirror kept faithful and kept current, and watched closely for the moment it starts giving orders. This is what it looks like to build one, accept that it can never be perfect, keep it the best available and endlessly improvable, and then, deliberately, step back and let the field hold it.
← Start the series
The Loop No One Chose
Reflexive dynamics in decentralized nonprofit fields, and why AI safety needs a live, actionable map that reflects the field back to itself instead of telling it what to do.
Also in this series
Open in Principle, Blind in Practice
Why not just fund the neglected directions? A funder willing to back overlooked work still needs the map — and without it reproduces the very loop they mean to escape, only running it in reverse.