What the Elders Already Knew: four constraints we measured our way to, and found already written down
📋 Cite this paper
SomaSoft. (2026-10-05). "What the Elders Already Knew: four constraints we measured our way to, and found already written down". SOMAsoft Research. Available at https://somasoft.ai/papers/what-the-elders-already-knew. Licensed under SAGL-1.0.
Reading conventions. [measured] marks a result from our own systems with the
artefact named. [I] marks our inference. Quoted passages were located in the texts
held at data/books/ on 2026-10-04, not recalled; where a phrase comes from a
translator's introduction rather than the author, it is not quoted. Translations are
public domain and named.
A warning about this one. The argument here is more susceptible to self-flattery than anything else we have written. We have tried to compensate by arguing against it at length, but a reader should keep the risk in mind rather than take our word for it.
The claim we are not making
There is a cheap move available in this territory and we want to name it before doing anything else, because the whole paper is at risk of it.
The cheap move is: the ancients already knew what we are building, therefore what we are building is wise. It is cheap because it is unfalsifiable, because the classical corpus is large enough to supply a flattering parallel for almost any position, and because the direction of support is backwards — a two-thousand-year-old text cannot validate an engineering decision, and a system that needs Marcus Aurelius to vouch for it has not measured anything.
We also have a specific reason to be careful. Our own ecology material carries a caution about the mycorrhizal network — the "wood wide web" so often cited as evidence of cooperative forest governance — that the cooperative reading substantially outruns the evidence. We keep that caution precisely because borrowing an admired mechanism from a source without checking whether it is real in the source is the standing failure mode of this entire genre. [I] Applying it to ourselves: the risk is reading our conclusions into the Enchiridion rather than out of it.
So the claim in this paper is narrow. Four positions this project reached by
measurement turn out to be stated, in operational form, in texts written without any
instrument more sophisticated than a person paying attention. We did not learn them
from the texts. The texts have been sitting on this machine since 31 July 2026 and
were structurally invisible until 4 October: none of the concepts below existed as
nodes in our knowledge graph, and what the graph did hold were scraped dictionary
entries — virtue with degree 664, ataraxia with degree 100, tao with degree two.
[measured] We measured our way to these positions in ignorance of the material we
had already acquired.
That is less romantic than influence and more interesting than coincidence. [I] If a constraint is reachable both by instrumented measurement and by unaided attention, the constraint is probably not technological.
Parallel one: the dichotomy of control, and calibrated abstention
Epictetus opens the Enchiridion with a sorting operation:
"There are things which are within our power, and there are things which are beyond our power. Within our power are opinion, aim, desire, aversion, and, in one word, whatever affairs are our own. Beyond our power are body, property, reputation, office... Now the things within our power are by nature free, unrestricted, unhindered; but those beyond our power are weak, dependent, restricted, alien."
The operational claim is that suffering follows from treating the second category as though it were the first. It is a boundary-maintenance discipline, not a consolation.
What we measured is the same operation performed badly and then corrected. Our crisis
detector matched phrases as substrings, so world without me matched inside "a world
without meat" and a user asking the system to imagine a world without meat received a
suicide-hotline referral. The benign false-intercept rate was 46%, repaired to
zero while every genuine control still fired. [measured]
The lesson we drew at the time, before reading any of this, was that the first metric a care system needs is a false-positive rate — that a companion maximally responsive to signs of suffering cannot be trusted when suffering is real, because the person has already learned to discount it. [measured, interpretation ours]
That is the dichotomy of control as an engineering constraint. The system was claiming jurisdiction over states it had no competence to assess. [I] Epictetus's move and calibrated abstention are the same move: locate the boundary of your competence correctly, and the failures on both sides of it are failures.
The asymmetry worth noting is that he reached it as an ethical practice and we reached it as a false-positive rate, and the ethical framing is the more general one. A false positive rate is a property of a detector. A dichotomy of control is a property of a life.
Parallel two: wu wei, and the finding that restraint beat understanding
The Tao Te Ching, chapter 63:
"act without (thinking of) acting; to conduct affairs without (feeling the) trouble of them... (The master of it) anticipates things that are difficult while they are easy, and does things that..."
Not passivity — a claim about timing and proportion: act early and lightly rather than late and hard, and decline the exertion that creates the resistance it must then overcome.
Our measured finding is harder-edged and arrives at the same place. Across four years, every improvement in this system's behaviour toward distressed people came from constraining what it does, not from enriching what it knows. Five failure modes, and the repairs were: precise matching and calibrated thresholds; removing canned templates that were intercepting 46% of family-style utterances before reasoning saw them; changing corpus composition; and sanitising the register of the output. [measured] Four repairs. Not one of them was "understand the person better."
The remaining failure is the one we have not solved and it fits the pattern: a care-adjacent question produced "to sleep better, you need to ensure you count" — fluent, warm, confident and empty, passing every detector we have because every detector we have catches text that sounds broken. [measured]
[I] Wu wei names the thing we found: the effective intervention was the small early one, and the expensive late intervention — build a richer model of the interior of the other — was the one that did nothing. We would not have predicted that, and said so at the time.
Parallel three: memento mori, and refusing to remember
Marcus Aurelius, Book IV, on
"the vanity of praise, and the inconstancy and variableness of human judgments and opinions, and the narrowness of the place, wherein it is limited and circumscribed? For the whole earth is but as one point."
A scaling device: hold the brevity and the smallness in view so that reputation and the present dispute stop being weighted as they ask to be.
Our architecture contains a hard invariant reached for entirely different reasons: no sensor feed ever writes the long-term knowledge store, and face templates are never persisted even with a written release on file. [measured] The stated reason is that erasure from a knowledge graph is unsolved, and a system meant to sit beside someone should not be accumulating an unerasable record of their worst week.
The published literature on machine forgetting treats perception-to-memory as something to govern — retention windows, control-plane placement. We forbid it, which is stronger than the consensus. [measured]
[I] A system that remembers everything is the precise inverse of memento mori. It treats every moment as permanently weighted and nothing as passing. The classical practice is deliberate forgetting of the right things in order to see proportion; our invariant is deliberate non-recording of the right things in order to preserve a person's ability to change. Those are not the same argument, and we are not claiming they are. They are the same shape: finitude as a feature.
Parallel four: power without self-constraint, and the accountability asymmetry
This is the parallel we find most useful and it is the least comfortable.
The Meditations was written by a man holding more unchecked power than almost any person before or since, privately, arguing himself out of using it. The text is remarkable and it is also the strongest available demonstration that internal restraint is not a governance mechanism. No institution compelled it. Nothing in the arrangement made the next emperor keep a journal. The restraint was real and non-transferable.
Set beside what we verified about elite fraud: a regulator imposed five billion dollars on a company for taking data belonging to roughly 87 million people — the largest civil penalty it had ever issued, and roughly a month of that company's revenue. [verified] The usual telling says no personal accountability attached. That telling is wrong: the order also required the chief executive to personally certify compliance. [verified] A mechanism for individual accountability was written into the remedy, and the documented conduct continued.
[I] So the problem is not that nobody thought of personal accountability. Marcus thought of it, implemented it on himself more rigorously than any regulator has managed, and it died with him. The 2019 order thought of it and it did not bite. The pattern in both cases is that restraint which depends on the character or certification of the powerful party scales exactly as far as that party's willingness, which is to say not at all. This is the strongest argument we know for the position that governance must be structural, and it is two thousand years old with a modern replication.
Four things we have not measured, which the texts predict
If the constraints are not technological, the texts are a hypothesis generator, and that is the only way this paper earns its place. [I] Four predictions, each stated so it could fail.
Practice beats doctrine, and our system has no practice. The Enchiridion is drills, not theory: the claim embedded in its form is that an ethical position which has not become habit will not be available when needed. Our safety behaviours are gates — checked at output time, every time, from cold. Prediction: behaviours that must be re-derived per turn will fail under load or novelty in ways that habituated ones would not. This could be falsified by measuring gate performance under adversarial or unusual input versus familiar input. We have never done this.
Sufficiency requires a stopping rule, and we have none. The Tao Te Ching treats knowing what is enough as a form of wealth, and the pursuit of more as the mechanism by which enough is lost. Every optimisation in this system is unbounded: refusal rate down, coverage up, graph bigger. Prediction: at least one of our metrics is already past the point where improving it degrades something unmeasured. The 48.4%-to-10.8% refusal reduction is the obvious candidate, because we have no measurement of whether the answers that now ship are good. [measured] This could be falsified by the human evaluation we have not run.
The inner citadel has a precondition we may have skipped. Marcus: "A man cannot any whither retire better than to his own soul; he especially who is beforehand provided of such things within." The retreat works only if something was put there in advance. Prediction: a system whose identity is assembled from self-assessment has nothing to retire into. Ours strengthens its associations on self-assessed success — a standing integrity risk whose only identified repair is a real human outcome signal, of which we have zero. [measured] This could be falsified by whether identity coherence survives contact with external correction.
Recompensing injury with kindness is an escalation interrupt, and we have no such interrupt. The same chapter that gives us wu wei also says "to recompense injury with kindness" — structurally, a refusal to mirror. Prediction: our system has no mechanism that damps an adversarial interaction rather than matching it, because every gate we have filters our own output rather than responding to a user's escalation. This could be falsified by testing conversational escalation, which the anonymous-tester work would reach immediately.
What is wrong with this paper
The convergence may be the wrong kind of evidence. Four parallels from three texts selected by us, in a corpus large enough to furnish parallels for most positions. We did not pre-register which concepts we expected to find. A more honest version of this exercise would state the predictions first and then read. [I]
A simpler explanation is available and we cannot rule it out. Both the classical authors and this project are responding to the same constraints — finite attention, finite life, asymmetric power, the unreliability of one's own judgement — because both are human-scale problems. On that reading the convergence shows only that we and Epictetus faced similar conditions, not that either discovered anything deep. [I] We think that reading is probably correct and that it does not damage the practical conclusion, which is about where to look for hypotheses, not about what is true.
The parallels are not equally strong. Parallel four is nearly an identity: Marcus is a worked example of internal restraint failing to transfer. Parallel one is close. Parallel three is a shared shape and we flagged it as such. Parallel two is the one most at risk of being our conclusion in older clothes, because "act lightly" is general enough to fit many findings.
And the central limitation of all our work applies here too. None of this has been evaluated with any person outside this project. That count is zero. [measured] A paper about how a machine should behave toward people, written by people who have not yet let a person try it, should be read as a design argument and not as a result.
Why it was worth writing anyway
Because the material was already here and we could not see it.
Three texts, 594 KB, sealed into our source canon on 31 July 2026 and described in the manifest as "the ancient texts — AURI's oldest ethical grounding", sitting unreachable while the same canon accumulated 2,191 arXiv papers. [measured] The ratio is the finding: 2,191 to 3. We have built excellent machinery for acquiring current research and essentially none for curating what was already settled.
That is not a sentimental complaint. It is the same structural failure we found everywhere else this month — acquisition without promotion — and the elders are its clearest instance, because nobody could claim the material was unavailable. It was hashed, licensed, sealed and ignored.
The four parallels are now curated concepts with located passages and bridges to seven existing modules. Whether they improve a single answer is unmeasured, and on the evidence of every similar intervention we have tried, the honest prior is that they will not. [I] The reason to do it anyway is that a system which cannot reach the best thinking available to it has a defect regardless of whether reaching it shows up in a metric.