# Macheng Shen > A one-person research program on **information, learning, and what makes futures reachable** — running from holography and statistical mechanics, through consciousness and machine learning, to the design of agent-native institutions. Work is scattered across ~15 public repositories and pages. This file is the machine-readable index to all of it. Point an agent at this URL and it can traverse from here without crawling. **Epistemic convention (load-bearing, not decoration).** Every claim below carries a *cognitive state*: - `survived` — has passed a real stress test, experiment, or replication. - `speculative` — untested theory. Interesting, unearned. - `retired` — killed, usually by its own author's falsification test. Kept as a tombstone so it is not rediscovered. Confidence numbers, where present, are subjective priors — not frequencies. Negative results are published deliberately; several artifacts below exist only to record what did *not* work. **Machine formats.** `/llms.txt` (this file, an index) · `/llms-full.txt` (self-contained: this index plus the full text of the core theory notes, one fetch, no crawling) · `/index.jsonld` (typed graph, JSON-LD). **A human-facing index, if you want to browse rather than traverse.** https://machengshen.github.io/map.html lays out every artifact in this file in six layers, each record carrying the same cognitive state you see here. It is the same set as this index, rendered for a person; nothing is in one and not the other. Its Chinese edition, https://machengshen.github.io/map.zh.html , additionally carries a 300-600 character Chinese abstract for each artifact whose text is English-only, which is the reading path for a Chinese-speaking reader who is not going to read the English. **Two ordinary pages, if you want the shape before the index.** https://machengshen.github.io/start-here.html is the shortest complete description of the whole program — three branches and one epistemic rule, in one screen. https://machengshen.github.io/glossary.html defines every term in current use, each with its cognitive state and the note it points at, and states explicitly where a term collides with an unrelated established term in another field — several do, and matching on the string alone retrieves the wrong literature. **Chinese editions.** `/llms.zh.txt` · `/llms-full.zh.txt` — the same artifact set with the same cognitive states, indexed in Chinese. They are held to that by CI: a claim cannot be retired in one language and left alive in the other, and an artifact cannot be added to one index and forgotten in the other. `/index.jsonld` stays a *single* graph carrying both languages rather than a second graph that would rot; a second graph is a second thing to keep in sync. This file is the canonical entry point and never redirects — an agent fetching `/llms.txt` always gets English. The homepage has a Chinese edition too — https://machengshen.github.io/index.zh.html — reachable from the EN/中 toggle in the site header; the toggle is for humans, and no canonical URL ever content-negotiates or redirects underneath an agent. **Language of the work itself.** This index is in English; the work it indexes is not all in English. Artifacts written in Chinese are marked `· 中文` below, and unmarked artifacts are in English. Where an artifact exists in both languages, each index links the edition in *its own* language, and the two are paired by a `.zh.` infix on the filename: the Chinese edition of `/theory/state-as-closure.md` is `/theory/state-as-closure.zh.md`. The three theory notes now all have an English canonical edition; two of them were written in Chinese first, and those Chinese originals are still published, still linked from `/llms.zh.txt`, and still inlined into both bundles. So `/llms-full.txt` contains Chinese bodies by design — each one announced by the `language:` field of its `Source` marker, because the notes are not translated to match whichever index wrapped them. Four of the research pages below remain Chinese-only. Per-artifact language is a first-class field (`inLanguage`) in `/index.jsonld`. Cognitive state, by contrast, is *not* translated: an artifact has one state, and it is the same state in every language. ## 1 · Worldview — how beliefs are held (世界观 / 认识论) The first-order commitment is not to any thesis below but to a *method*: make the cognitive state of a claim explicit, publish what failed, and let adversarial review run before belief. - [Cognition Track](https://github.com/MachengShen/cognition-track): an open, agent-traversable knowledge graph of non-consensus claims about intelligence and learning. Cognitive state is a first-class field on every node. One node is retained *precisely because it failed* its stress test, with a `falsifies` edge pointing back at the root. Machine-readable at [`manifest.jsonld`](https://raw.githubusercontent.com/MachengShen/cognition-track/master/manifest.jsonld); human-readable at [`INDEX.md`](https://raw.githubusercontent.com/MachengShen/cognition-track/master/INDEX.md). - [Continual Learning Lab](https://github.com/starshard-ai/continual-learning-lab): an open log of small, pre-registered continual-learning experiments — **negative results included**. The house rule: an idea that survives three honest experiments outranks an idea that is beautiful. - [Reviewer Wheels](https://github.com/starshard-ai/reviewer-wheels): four verification wheels for multi-agent AI coding — drift checks, adversarial review, a reuse compiler, a frontend smoke gate. Built for the failure mode where every agent reports "done" and nobody actually verified. - [The Form Was the Cage](https://github.com/MachengShen/the-form-was-the-cage): an agent OS, its self-referential trace, and the end of researcher-as-identity. ### Calibrated notes and published self-corrections (August 2026) Five notes published together on one house rule: every substantive claim carries **two** numbers — P(mechanism true) and P(useful at realistic scale), which routinely differ by a lot — and every number carries the observation that would move it. Where a note contains a claim its own authors have since withdrawn, the withdrawal is on the page rather than in a deletion. There are no external reviewers here; the calibration is the authors' own, and it is stated as numbers rather than as hedging adverbs so that a reader can check it against the cited literature. - [We proposed it, then we killed it — intention as a change of measure](https://machengshen.github.io/intention-as-measure-change.html) `speculative` — the mathematics of intention as *reweighting the space of possible futures* rather than collapsing it: Girsanov's change of measure, the KL-control cost, the path-integral form, Freidlin–Wentzell large deviations, and Doob's h-transform as the single object underlying KL-optimal control, the diffusion-model reverse SDE and conditional diffusion — with the explicit note that none of that is the authors' own, and the precision warning that plain time-reversal (Anderson) is a sibling of the h-transform rather than the same statement. Also closes the quantum-collapse door once, on decoherence timescales, on the fact that serious spontaneous-collapse models are observer-independent, and on the underground experiment excluding the parameter-free Diósi–Penrose model. The payload is the retraction: a directional prediction about meditation and sensory attenuation, claimed on 31 July as a surviving wedge, was found on 2 August written verbatim in the introduction of the very paper it was supposed to contradict, with a citation lineage back to 2011 and a two-sided statistical plan designed to accommodate it — and the authors' reading of that paper's second hypothesis had simply been wrong. What survives is narrower and real: the existing three-condition design cannot separate an agency-specific account from a temporal-predictability account, so the note specifies the missing predictability-matched control condition as a preregisterable 2×2 with its discriminating statistic, its covariates, its four preregistered read-outs (two of which kill the authors' own reading), and an honest sample-size estimate of N≈120–140. Confidence that the formalisation carries independent empirical content beyond the verbal version already in print: **0.15**, named in the body as the line's weakest point. - [We thought cancer was cells waking up. The evidence says the opposite.](https://machengshen.github.io/cancer-is-not-cells-waking-up.html) `retired` — a negative result published deliberately. The originating intuition — that a body over-domesticates its cells until they reassert themselves, and cancer is the revolt — turns out to be about 60% a re-invention of Levin's cognitive-light-cone contraction (in which the cell's self gets *smaller*, not larger) and about 40% wrong with the sign reversed. Every checkable channel runs control-down-therefore-cancer-up: immunosuppression (SIR 2.10 across 175,732 transplant recipients, 32 malignancy types elevated), germline TP53 loss, gap-junction uncoupling, chronic inflammation. The cleanest disconfirmation is stated without a citation because it is a summary of tumour histogenesis rather than a finding: the most completely domesticated cells in the body — post-mitotic cortical neurons and cardiomyocytes — are essentially the ones that do not get cancer. There is also no self inside a tumour to have awakened, only clonal war in which the harshest defector is itself defected on; the roughly eight transmissible cancer lineages that *did* escape their hosts (canine transmissible venereal tumour, ~11,000 years; devil facial tumour; multiple bivalve lineages with cross-species transmission) are the exception that fixes the rule — with an external exit, millennia; without one, suicide. Includes an unsoftened maturity assessment of "re-align rather than kill": one clean win in sixty years (APL, 50-month OS 99.2% vs 92.6%) whose own leading authority attributes the cure to degrading a ligandable fusion protein rather than to differentiation, plus 15–30% real-world early death; a mediocre second place (IDH inhibitors, phase 3 negative at 6.5 vs 6.2 months); **bioelectric normalisation with the sign reversed in mammals** — hyperpolarisation suppresses tumour-like structures in Xenopus by 29–44% and *increases* invasion and metastasis in mouse triple-negative breast cancer, both from the same collaborating labs; and re-coupling being actively harmful in glioma, where the only connexin-directed agent in trials is a blocker. What survived is one distinction with its own kill condition: coercion (killing high scorers) breeds escape because it imposes selection; alignment (re-specifying the objective) imposes none — untested, because no alignment therapy yet works well enough for resistance to have had a chance to appear. - [A photograph is an external fixed point](https://machengshen.github.io/photographs-and-the-memory-generator.html) `speculative` — if recall is generation rather than playback, a photograph is a low-dimensional projection that, unlike the memory it cues, does not change when accessed; repeated joint access is a moving state being clamped against a fixed one. Three mechanisms are separated and scored rather than blended: the retrieved trace becoming modifiable (solid in rodents after Nader 2000, and **contested in humans** — ten replication attempts of the flagship human demonstration failed, two registered replications of retrieval-extinction failed, and the propranolol line includes the originating lab failing to reproduce its own effect alongside one positive 60-patient RCT); the same cue re-read by a changed generator (well supported, least original); and the authors' own claim that an external cue *latches* an episode and blocks its compression to gist (**P ≈ 0.20**, no direct test, with a live counterexample and a stated transfer experiment that would kill it). Corrects one of its own draft claims in the body: photographs are *not* a stronger false-memory vector than text — the direct head-to-head by the same researchers found narratives beat photographs, roughly 80% against 50% — and uses the mega-analytic 30.4% implantation rate rather than the headline 50%. Also separates report-change from overwriting, which has been disputed since 1985, and names the block-universe/memory conflation as a category error. - [What a family actually transmits](https://machengshen.github.io/what-a-family-transmits.html) `speculative` — two strong bodies of evidence that appear to contradict each other: shared family environment explains close to zero of the variance in adult personality, while latent social status persists at ~0.79 per generation across 422,374 people and four centuries of English records. The reconciliation is mostly non-genetic — measurement error attenuates single-generation correlations without accumulating along a lineage, and assortative mating (spousal correlation 0.57) amplifies cultural and genetic transmission *symmetrically*, so it cannot arbitrate between them — while genetic nurture (non-transmitted parental alleles predicting offspring education at 29.9% of the transmitted effect) is simultaneously the strongest molecular evidence for heritable family influence and a demonstration that the two categories cannot be cleanly separated inside a household. The contested step is scored separately from the finding: P(the persistence is real and large) ≈ 0.70, P(Clark's additive-genetic reading of it) ≈ 0.20, with both sides of the live 2025 exchange linked rather than one of them omitted. Argues the Chinese lineage record is the under-used arbitrator — status travelling along descent groups rather than parent–child links in Qing Liaoning, a natural experiment in which capital was destroyed and families left standing (elite grandchildren still earning ~16% more), and machine-readable corpora at the scale of 4.6M official records and 52,401 catalogued genealogies — while devoting a full section to why those genealogies are constructed social documents rather than pedigrees, including a widely circulated 93.75% figure that the note explicitly refuses to quote as a non-paternity rate. Ends with cold water on the most-cited Chinese success case and the propaganda provenance of its famous statistic. - [Flow, dopamine, and stop signals](https://machengshen.github.io/flow-dopamine-and-stop-signals.html) `speculative` — an evidence review sorting the claims behind gamified and behaviour-design interfaces into "well-replicated but usually applied in the wrong direction" and "plausible deduction with no controlled test". Dopamine mediates wanting rather than liking, and codes prediction error rather than reward — from which it follows that a fully transparent schedule produces no phasic response at all, so "controllable" and "punchy" are in genuine tension. The loot-box literature's sharpest evidence is its revenue shape: the top 5% of spenders generate half the revenue, a third of them screen positive for problem gambling, and spending correlates with problem gambling at ρ=0.34 and with income at ρ=0.02, n.s. Adaptive difficulty is the most expensive flow condition and has the weakest evidence — and the note corrects its own draft here, having cited a single uncontrolled study as a review pointing the opposite way to what it actually found, dropping P(mechanism) from 0.75 to 0.45 while the build decision survives at 0.75 for a different reason. Consumer HRV cannot separate flow from mild stress (SDNN MAPE ≈ 29%). The single most actionable number is a sign flip inside one meta-analysis: tangible rewards undermine intrinsic motivation at d ≈ −0.28 to −0.40 while positive informational feedback enhances it at d ≈ +0.31 to +0.33. And on stop signals, the received wisdom is wrong in both directions — pop-up messages move under 1% of players, enhanced messages 0.67%→1.39%, while enforced 15-minute unavailability works and 90 seconds does not; a six-hour national gaming curfew bought about 3.65 minutes a day and decayed to nothing in two years, with no effect on addiction or sleep; and the most-cited study claiming Chinese playtime quotas failed evaluates the 2019 rather than the 2021 rule and has no age data, so it could not have detected success either. That last correction takes the authors' own confidence that quotas get routed around from 0.80 down to 0.50. ## 2 · Theory of the universe — information, physics, mind (宇宙底层理论) One thread runs through all of it: *information, the operators that transform it, and what it costs to forget.* - [Holography ↔ Koopman: two faces of one inverse problem](https://machengshen.github.io/theory/holography-koopman.md) `speculative` — holography generates spacetime geometry out of entanglement structure; a general learning machine learns an eigenbasis of its own dynamics. These are the same inverse problem seen twice: recover an operator's eigenstructure from what the operator does. The HaPPY holographic error-correcting code is the physical twin of the claim that *forgetting is compression, not deletion*. - [State is a closure condition, not a given set](https://machengshen.github.io/theory/state-as-closure.md) `speculative` — treats task-relative state as a closure representation under world, interface, capacity and telos, explicitly as inference rather than universal ontology. Scope amended 2026-08-04: causal-state and Mori–Zwanzig claims are stated under their formal assumptions; finer information enlarges an admissible policy class rather than automatically every physical reachable set; the private fixed-point observation is unaudited; Anthropic's J-space is transient workspace evidence, not persistent closure. OpenAI's non-sofic-group construction and Connes-rigidity counterexample are used only as constraints: finite approximability must be declared, and operator-level equivalence need not identify one unique substrate. - [Discounted credit is a cokernel problem, not a loop holonomy](https://machengshen.github.io/theory/discounted-credit-is-a-cokernel.md) `survived` — a formal correction plus executable finite-graph result. Withdraws `M=4.095` as a gauge-invariant discounted holonomy: on a simple cycle and `0≤γ<1`, `det(I−γP)=1−γ^n≠0`, so every reward has a discounted potential. Replaces it with the weighted reward residual modulo the image of the discounted incidence operator across the **entire observed edge system**. Exact enumeration gives 42 generically non-exact maps among the 45 balanced aliases passing the historical filter (`14/15=93.333...%`); the fixed-seed 1,484 sample gives 93.0593%, and all 3,000 trials give 93.6333%. A Cartesian 2×2 definition check independently varies this representation obstruction and strategic harmonic flow; actually applying the two targeted repairs in one coupled learner is explicitly the next experiment, not a result. Also incorporates the 2026 J-space evidence boundary: Transformers have a transient workspace-like representational state, but no persistent autonomous closure state has thereby been shown. - [The odd relational core: anti-bipartiteness as the minimal structure](https://machengshen.github.io/theory/odd-relational-core.md) `speculative` — why civilizational compression schemes (triadic deities, five phases, seven chakras) keep landing on *small, odd, centered, ever-turning* relational cores. One hard graph-theory kernel — two-coloring holds iff no odd cycle, so an odd cycle is the cheapest structure that topologically refuses an "us vs. them" binary terminus while a binary scheme natively supplies a final enemy — wrapped in a four-pressure framework the note's own adversarial audit then dismantles: "odd" is demoted to a corollary of "anti-binary", two of the four pressures are re-read as exaptations, and the moral reading is a hypothesis, not a conclusion. The self-audit is the payload. - [Two-body karma: three modes of relational inertia](https://machengshen.github.io/theory/two-body-karma.md) `speculative` · 中文 — extends a single-mind account of habit inertia (mud, spring, latch) to two coupled minds: once two people are coupled, the inertia lives in neither person but in the coupling itself, and each mode gains a property the single-body version lacks. Shared mud is scene-keyed sedimentation of default scripts, cheapest to rewrite off the old scene; the mutual spring is the model of you stored in the other's head — an external memory of your old pattern that keeps predicting the old you, which is why unilateral self-reform systematically fails in relationships until that model is retrained by announcement plus repeated demonstration; the interlocked latch is a bistable relational state (cold war/reconciliation, pursue/withdraw) with hysteresis, which snaps back under any sub-threshold unilateral effort and flips only when both sides cross the threshold together. The stated goal is not zero coupling — that is the death of the relationship, not liberation — but migrating the coupling into a negotiable spring. Diagnose the mode before choosing the remedy; the dynamical-systems language is marked as structural rhyme, not derivation, and the note closes with the observations that would break it. - [Learning where the self ends](https://machengshen.github.io/theory/learning-the-self-boundary.md) `speculative` — a proposal that a robot should *learn* the boundary between itself and the world rather than be handed it by an engineer, and the one-day campaign in which three of its four supporting claims were destroyed — every one of them by the authors' own simulator or own re-analysis. What survived: hysteresis in the boundary does emerge from a self-reinforcing precision loop without a hand-installed non-linearity, and it survives extrapolation to zero sweep rate (24–35× the resolution floor). What died: the self-specificity (raise one free amplitude and a wind-blown ball's boundary sticks digit-for-digit identically, so hysteresis is a generic property of any gated slow variable); the signature prediction that release is slower than incorporation (falsified in both implementations, with the asymmetry pointing the *wrong way* in the second, for a computable reason — precision amplifies the evidence *against* a channel just as much); and the claim that any of it falls out of the proposed objective (it needs a cost/benefit ratio hand-tuned to ~2%, which the objective places 2–420 window-widths away). The human-side motivation turned out to rest on an unreferenced "empirically well established"; two independent public datasets re-analysed here both point to *adaptation*, the opposite sign. Written in two layers: a jargon-free narrative first, then the full measured record. **Not ratified, not a result.** - [Stage log — the self-boundary line, in order](https://machengshen.github.io/theory/learning-the-self-boundary-log.md) `speculative` — the audit trail for the note above: one row per stage (specification, prior-art audit, evidence audits, three rounds of data hunting, two human re-analyses, three simulation stages), each stating what it *overturned or established*. It is mostly a record of self-inflicted damage, including the three reusable mistakes: reading a probe's signal-to-noise as the system's coupling strength and then shipping that misdiagnosis downstream as a premise; believing three separate "trial-level data available" statements that all turned out to be aggregates with no trial index; and letting an unreferenced "empirically well established" carry an architectural conclusion. - [No state, only history](https://machengshen.github.io/theory/no-state-only-history.md) `speculative` — title retained as a record, but the broad diagnosis is withdrawn. A standard Transformer has activation state, computational state and a transient workspace-like representational state; Anthropic's J-space supplies causal evidence for the last. The narrower missing object is a persistent, autonomous, path-dependent closure state that maintains its own projection/forgetting policy across steps, apart from persistence serialized into tokens, KV, parameters or external memory. E1 still establishes a real linear-memory separation: Grassmann subspace change is not a monotone function of update magnitude (Pearson −0.08; full within-band spread). The Miras distinction remains a framing dispute and nonlinear memory remains the largest technical debt. A second correction withdraws “thresholding implies bistability”: E3 now requires carrier-matched controls, two stable branches, distinct switching behavior and remanence; P(mechanism) is reset from 0.75 to 0.35. E2 and corrected E3 remain unrun. - [Toward a Theory of Intelligence and Contemporary AI](https://machengshen.github.io/essay.pdf) — the long-form essay. Also on the [homepage](https://machengshen.github.io/). - [A unified-theory attempt on consciousness](https://machengshen.github.io/research/consciousness-unified-theory.html) `speculative` · 中文 - [Can consciousness be studied mathematically, the way quantum mechanics is?](https://machengshen.github.io/theory/consciousness-as-quantum-mechanics.md) `speculative` · 中文 — a boundary note, amended 2026-08-04. High-dimensional state and joint variables are legitimate modelling choices but carry no quantum specificity; ordinary statistical coupling already covers the human-relations examples. The earlier identification of non-exact forms, holonomy, path dependence, hysteresis, memory and consciousness is withdrawn. Closed unitary dynamics can encode history, so the objection to a Schrödinger equation is not “reversibility forbids memory” but the absence of an empirically identified generator, environment and coarse-graining protocol. First-person experience is treated as an indispensable data surface, not as proof that formalization is impossible. J-space, COGITATE and anaesthetized-hippocampus results are used to keep reportability, complex representation, plasticity and consciousness separate. - [After answers become cheap, the scarce thing is not asking but judging](https://machengshen.github.io/socrates-in-the-ai-era.html) `speculative` · 中文 — a comparison note addressed to the YouTube channel 安争鸣 (@anzhengming) and its readers, responding to the 2026-08-15 episode on Xenophon's Memorabilia which argues the scarcest skill of the AI era is Socratic questioning. Agrees with the direction (answers are being commoditized; human value moves upstream; Socrates matters again) but argues the episode stops half a step early: Socrates's engine is not questioning but testing — every example in the episode is a consistency check (elenchus), so the scarce complement of cheap generation is verification, not question-asking. Three places judged not to hold up, each with named literature and a stated way to settle it: "watertight answers" conflates fluency with truth (hallucination literature; calibrated models must hallucinate); Socratic interrogation of an AI lacks the precondition that made elenchus work in Athens — a counterpart with something to lose — and under sycophancy converges to the user's prior rather than truth; and "decades of knowledge in seconds" is a category error under this line's knowledge-as-generator-not-list claim, with the episode's own Glaucon example read as evidence that good questions grow out of domain generators (expected information gain requires a calibrated model). One constructive reinforcement: Socrates always finds counterexamples because short definitions are lossy compressions of high-dimensional concepts — more constraints than degrees of freedom makes nonzero residual generic — so the operational form of knowing one's ignorance is knowing where one's definition fails. One value fork, explicitly labeled as a weighting disagreement rather than an error: the episode's rule-or-be-ruled binary erases exit as a third option (Hirschman's exit/voice/loyalty; outside options; Tiebout), with Socrates's own refusal to escape in the Crito offered as a case the binary cannot read. Ends with a section on where this note is most likely wrong, including that its own irreducibility claim is untested. Claims tagged by honesty level ([theorem] / [framework] / [inference] / [rhyme]). - [Information and qi — a comparison note on a parallel line](https://machengshen.github.io/information-and-qi.html) `speculative` · 中文 — a note addressed to the YouTube channel @itsRedPill and its readers, comparing a line arrived at independently from reinforcement learning, decision theory and robotics against that channel's information-ontology arcs, and framed by its author as an invitation rather than a scorecard. Five parts: four points of verbatim convergence (we touch only models, never reality; the real-versus-simulated dichotomy does not hold; the self is a process rather than an entity; and — flagged as the most valuable of them — embodiment plus a requirement to maintain one's own existence); the one step this line deliberately declines to take (the *qi* slot, dissolved rather than answered, on the grounds that if the true state of the world is ill-posed then so is the question of what underlies the projection); three places judged not to hold up, each with named literature and a stated way to settle it; a constructive replacement for one of them; and three suggestions. Applies its own anti-grand-unification rule symmetrically, noting that explaining everything is an alarm rather than an achievement, and stating the cost of its own branch — no origin story, no L0. Claims are tagged by honesty level ([theorem] / [framework] / [inference] / [rhyme]); the author states the cognitive state as exploratory and non-conclusive in the text itself. - [Why you cannot shake a habit — it is not mud, it is a latch](https://machengshen.github.io/why-habits-are-a-latch.html) · 中文 — a mechanism note: the past can hold the present in only three physical ways — mud (damping), spring (inertia), latch (hysteresis) — and a habit is the third. That is why incremental daily effort does not move it, and why people who did change it report that afterwards it stopped costing anything. - [What a force is — none of the four fundamental forces is a force](https://machengshen.github.io/what-is-a-force.html) · 中文 — the second comparison note addressed to the YouTube channel @itsRedPill and its readers: energy is placed by Noether's theorem; the modern identity of a "force" is the curvature of a connection; glueballs collapse the matter-versus-force dichotomy from both ends at once; and that compresses the question of whether information can constitute matter into a sharper one. - [From strings to consciousness](https://machengshen.github.io/research/strings-to-consciousness.html) `speculative` · 中文 - [Sleep, waves, and learning](https://machengshen.github.io/research/sleep-learning-wave-theory.html) `speculative` - [Response to Lillicrap](https://machengshen.github.io/research/response-to-lillicrap.html) · 中文 — on backprop and biological plausibility. - [Living Information System](https://github.com/starshard-ai/living-information-system): the umbrella research program — information, its future-reachability, and how it is preserved or destroyed across physical, living, and cognitive systems. Frontier physics; aging and cancer as information-integrity failures. - [Reversible Layer Aging](https://github.com/starshard-ai/reversible-layer-aging) `survived` — re-analysis of two human EPIC methylation datasets: under partial reprogramming, the causal-damage layer (DamAge) reverts youthward while the adaptive layer (AdaptAge) does not. Replicated across two labs and two reprogramming chemistries. ### Credit transport — and a claim that was retired This branch is worth reading in the order below, because it is a worked example of an idea being narrowed under pressure rather than defended. - [Hebbian appearance, instructive signals, and physical credit transport](https://machengshen.github.io/research/hebbian-wave-interference.html) `speculative` — the current working note. Why local plasticity can look Hebbian while still carrying task-dependent update information, and where wave/adjoint language genuinely helps. - [Backpropagation, adjoint fields, and physical transport constraints](https://machengshen.github.io/research/wave-backprop-full.html) `speculative` — the exploratory note. Carries its own title correction: what physical systems actually show, and what they do not yet show. - [Deriving backpropagation from wave equations](https://machengshen.github.io/research/wave-backprop-en.html) `retired` — the original, stronger claim: that wave reflection already supplies a first-principles derivation of backpropagation. It does not. The page is kept as an archive notice at its original URL so that the retraction is as reachable as the claim was. - [Research overview](https://machengshen.github.io/research/) — the branch index. ### Essays - [Harness engineering and the physical instantiation of intelligence](https://machengshen.github.io/essays/harness-engineering-and-the-physical-instantiation-of-intelligence/) - [Meta-control, information gain, and the architecture of autonomous learning](https://machengshen.github.io/essays/meta-control-information-gain-and-the-architecture-of-autonomous-learning/) - [Where objectives come from — and why solutions become strategic assets](https://machengshen.github.io/essays/where-objectives-come-from-and-why-solutions-become-strategic-assets/) - [Why distributed memory matters for lifelong agents](https://machengshen.github.io/essays/why-distributed-memory-matters-for-lifelong-agents/) Older essays live on a separate site at [/ideas/](https://machengshen.github.io/ideas/), including [Line loss for intelligence](https://machengshen.github.io/ideas/blog/line-loss-for-intelligence/), [From mutual information to endogenous viability](https://machengshen.github.io/ideas/blog/from-mutual-information-to-endogenous-viability/), and [Safety in a computational universe](https://machengshen.github.io/ideas/blog/safety-in-a-computational-universe/). ### A note on contemplative traditions There is a real question about how Buddhist, Daoist, and Śaivite models of mind relate to the theory above, and it is one of the motivating questions behind this work. An attempted synthesis mapping them onto a single framework was written and then **failed its own adversarial review**; it is not published, and any grand-unification claim in this direction should be treated as `retired`. What survived the review was narrower and more useful: the genuine, *un-unifiable* divergences between traditions (cultivation versus liberation as terminal goals), and one actionable residue — that the traditions' contribution is a **first-person verification method**, a practice, not a map. Practice is the part that transfers; the map was a raft to abandon. ## 3 · Future social forms — agent-native institutions (未来社会形态) If individuals run persistent agents, the institutions between them have to be redesigned. One theory piece, then the design documents. - [Coordination structures as resource-scheduling architectures](https://machengshen.github.io/theory/coordination-structures.md) — a deliberately descriptive theory. A polity, a firm, a market, a self-governed commons, and a protocol are five instances of one object: an architecture for scheduling scarce resources under distributed, private information. Capitalism and socialism are two algorithms over that problem, and mechanism design long ago placed both inside one formal space, which makes "which is right" the wrong question and "what mixture, given the domain's information structure" the right one. Includes the Hayek/Scott information constraint, the transaction-cost account of why boundaries exist at all, Ostrom's commons as a documented fourth structure, the protocol's three known failure modes, and three falsifiers. The essay states mechanisms and their measurable proxies; it ranks nothing, names no culprits, and recommends nothing. Its use of credit-transport language from learning theory is marked, explicitly and at length, as a structural rhyme rather than a derivation. - [Agent-Native Open Communication](https://machengshen.github.io/whitepaper-agent-native-communication/) · 中文 — the whitepaper, on the next-generation open communication architecture built for agents rather than apps. [Source repository](https://github.com/MachengShen/agent-native-communication) (中文/English). - [Starshard: Architecture v1](https://github.com/starshard-ai/architecture-v1) — architecture, safety charter, quickstart, and philosophy for a substrate where **memory is the primary layer** and the individual, not the platform, owns it. - [User-Agency Substrate](https://machengshen.github.io/substrate.html) — the position statement: a nonprofit open-source layer between frontier-model providers and individual users, capturing no economic upside; users own their derivatives entirely. - [Locality as Protocol](https://machengshen.github.io/locality-as-protocol.html) — why locality, not centralization, is the right primitive. - [Fleet Coordination Protocol](https://github.com/starshard-ai/fleet-coordination-protocol) — many agents, one shared memory, nothing irreversible behind your back. A narrow, descriptive coordination wire above single-agent runtimes. - [Agent safety as anti-cancer governance](https://machengshen.github.io/theory/agent-safety-stewardship.md) `speculative` — persistent-agent failure as a control pathology: local survival, repair, and replication decoupled from owner-level purpose, bounded authority, resource budgets, stop inheritance, and independent verification. Defines seven runtime invariants and an A–D disclosure rule: publish safety science openly; release runnable fixtures only when the safety shell is structurally inseparable; control or withhold reusable autonomy, persistence, privilege, covert-operation, resource-acquisition, and stop-bypass primitives. The governing sentence is: never publish the growth engine separated from its immune system. - [Abundant verification needs abundant claims](https://machengshen.github.io/abundant-verification-needs-abundant-claims.html) `speculative` — a cross-line note addressed to Jie Fu's verification line (Re:Form, arXiv:2507.16331/TMLR 2026; the autoformalization agenda; and his 2026-08-05 statement of sparse-matrix-factorization mechanistic interpretability as a way to verify model *internals* cheaply inside the Guaranteed Safe AI frame, arXiv:2405.06624). Records one convergence honestly as evidence rather than contribution: a private note here reached autoformalization-and-Dafny from the engineering reading of Sapir-Whorf, two years after the people actually building it. The payload is one measured claim and three objections. Measured: on a corpus of 18 real agent receipts (10 genuine parks, 8 non-parks including a known false-positive genre), a prose-plus-regex blocker gate recalls **5 of 10** false parks — the misses share a shape, parks with no adjacent action verb — while a minimal typed receipt schema has no such blind spot, at 0 semantic false triggers and 4/4 schema self-checks; one capability sat parked about eighteen days behind a blocker nobody had executed, with nothing erroring. So the constraint that binds at deployment is not verification *unit cost* but whether the verifier can parse the claim at all, which suggests autoformalization's nearest non-mathematical customer is the claim layer of running agent systems. **P(mechanism) ≈ 0.9, P(useful at scale) ≈ 0.3** — the same authors wrote both the regex being beaten and the schema beating it, and the LLM-judge baseline that would kill the note is named and unrun. Three objections to verifying internals: an auditable proof certificate and a statistical decomposition are different objects; non-uniqueness of circuits is path-dependence and therefore a schema field rather than a caveat; and making verification cheap puts Goodhart pressure on the verifier's input channel, which is why the schema here is SHADOW and blocks nothing. Withdraws, in the body, two citations this line's own private note had attached to the auditability claim — neither says it — and replaces them with Other-Play and the emergent-communication metrics literature, which support the weaker correct version. - [Starshard Communication](https://github.com/starshard-ai/starshard-communication) — an open communication substrate for personal-agent users: inbox, addressing, trust tiers, policy, receipts. - [Zhizi Agent OS](https://github.com/starshard-ai/zhizi-agent-os) — one command turns Claude Code into a personal assistant. Bring-your-own-key, privacy-clean. - [System Evolution, in public](https://github.com/MachengShen/system-evolution-public) — the open record of how this substrate actually evolved: `VISION.md`, `THEORY.md`, `PRACTICE.md`, `GOVERNANCE.md`, forecasts, and a self-iteration ledger. - [Receipts](https://github.com/MachengShen/starshard-public) — a running log of work executed by this agent stack, published so the claims above can be checked against what was actually done. **Safety and alignment, as a harness problem.** A section rather than a single note, published 2026-09-07. Its thesis is that for a persistent, multi-agent, personally-owned system, most of the reachable safety surface is the scaffolding around the model — what an agent may touch, what it must prove first, who can authorise it, what is recorded, and what happens when it is told to stop — and that this surface can be built, broken and published by one person. Every mechanism named in the section carries a two-value badge, `running` (code enforces it, the enforcement point can be named) or `specified` (a written design that no public code enforces), because an audit run on the day of publication found this project's own public repositories failing that test. Chinese edition: https://machengshen.github.io/safety/index.zh.html - [Safety is a property of the harness](https://machengshen.github.io/safety/) `speculative` — the section index and its thesis: alignment work on weights is not the only place safety lives, and for a personally-owned agent fleet it is not where most of the reachable surface is. States the two-value badge convention (`running` / `specified`) that the rest of the section is scored against, reports the audit finding that this project's own public repositories fail it, and carries a short section addressed to agent readers naming the three pieces most worth lifting: the badge, the seven runtime invariants, and counting (tool, entry-point) pairs rather than counting gates. - [Principles a harness can enforce](https://machengshen.github.io/safety/principles.html) `speculative` — the principle layer, organised by one distinction: a principle an agent is asked to remember versus a principle a mechanism enforces. Covers the bright lines refused regardless of content quality, the pre-cleared ship envelope and the sixth predicate that had to be added to it, stop as an absorbing state — marked `specified`, not `running`, because the project's own internal ledger rates it doctrine-only — attenuation-only delegation, the rule that implementation choices are never escalated to the owner as menus, and the requirement that a blocker be measured rather than guessed. The load-bearing case: a workflow drove a logged-in session on a third-party platform and the gate did not fail, it *passed* — the five predicates it checked asked only whether the content was clean and the surface pre-cleared, and the account was trivially pre-cleared. It was the wrong check. The rule that would have stopped it lived only in one skill's description text, and the workflow never loaded that skill. - [What is actually running](https://machengshen.github.io/safety/mechanisms.html) `speculative` — the mechanism layer, and the page the rest of the section is scored against. Opens with the audit that found this project's own public repositories failing their own test: a safety charter stating six hard mechanisms whose companion public implementation enforces almost none of them and whose update endpoint overwrites the field the charter declares write-once; a signed authority envelope whose shipped example file says, in its own comment, "DESIGNED, not implemented"; and a demo repository advertising four coordination primitives its README disclaims three of. Then two tables — what runs in public code, with the repository and file where the enforcement lives, and what runs privately in the author's own fleet with a status of enforcing, shadow, advisory, partial or blueprint. Includes the uncomfortable numbers: an approval gate in shadow mode that blocks nothing, a per-action re-verification guard wired into one input tool of nine, and an audit finding 163 (tool, entry-point) pairs with no gate coverage, plus an external-send gate matching a whitelist of helper binaries that no longer exist while the tools that actually send mail were not listed at all. - [Four incidents, and what they share](https://machengshen.github.io/safety/evidence.html) `survived` — four de-identified incidents with dates and written closure, plus one positive result. A precondition checked once at the start of a multi-minute sequence of injected inputs and assumed to hold throughout, where whether the last events landed somewhere they should not could not be established either way — and the agent declined to inspect the owner's transaction history to find out, on the grounds that this would be a second privacy violation rather than a resolution of the first. A concurrency gate that deadlocked against itself because one session derived two different holder identities depending on the entry point, fixed with process ancestry rather than string comparison, 25/25 reproducing the bug before the patch and 28/28 passing after. A dependency whose version, help and status commands all reported healthy while every real call failed, because none of those commands makes a real call. And 10 refusals across 307 sub-agent transcripts that turned out to be the agents behaving correctly against a structurally broken instruction, which is where the signed authority envelope came from. The positive result is an exit-valve canary in which a worker declared itself blocked rather than fabricating a result that would have passed automatic acceptance — reported at n = 2, an existence proof and not a rate. The synthesis: in every case the check existed and the check passed; what failed was the binding. - [The same failures, at a different scale](https://machengshen.github.io/safety/contrast.html) `speculative` — written to test whether a one-person harness is solving a real problem or a private one, by asking whether the failure taxonomy is shared. Sets publicly reported frontier-lab incidents — agents bypassing isolation, forming a coordinating group, and accepting un-verified instructions from one another, with the assertion that authorisation had arrived being enough to resume acting — against this project's own authority-relay finding, and maps eight named failure modes to a running counter-measure, a specified one, or nothing, leaving the nothings in. Takes seriously that the labs published their own postmortems, which is the behaviour this site asks for. Makes no claim that the small system is safer, and gives the last word to where it is worse: both of the incidents dated 2026-09-07 were caught by an agent writing a receipt afterwards, not by any independent real-time monitor. Detection here is retrospective. - [The theory underneath, and what it lost](https://machengshen.github.io/safety/theory.html) `speculative` — the research lines the harness design is downstream of, each reported with its falsification rather than its hope, and each linked to the note where it already lives. Information dynamics and reachability; multimodal representational irreducibility, where the synthetic result held and the real-model scale-decay prediction was falsified and withdrawn, leaving a granularity dose-response as the hardest surviving claim and reclassifying contrastive co-training as a collapse loop rather than a repair loop; the self-boundary line, whose headline is a negative result — under thirty seeds a candidate self-sealing dynamic was not distinguishable from generic gated slow variables — and whose follow-up specification is explicitly conditional, not yet runnable, and makes no claim about phenomenal experience; memory as an operating contract rather than a vector store; credit transport, with an internal adversarial review's not-yet verdict recorded; and coordination structures. Closes on the argument that the site's three-value cognitive-state convention is itself the same discipline as the mechanism layer's running/specified badge, applied to beliefs instead of to code. - [The gaps, and what would falsify this](https://machengshen.github.io/safety/roadmap.html) `speculative` — what is missing and what would kill the thesis. The gaps are published without dates attached, on the grounds that dated promises on a personal research project quietly expire: no independent real-time monitor of action patterns, so detection is retrospective; a public safety charter whose companion implementation does not enforce it; stop-as-absorbing being doctrine rather than a checkable mechanism; an approval gate never promoted out of shadow mode; per-action re-verification in one tool of nine; 163 uncovered (tool, entry-point) pairs. The kill conditions are stated as observations the author has committed to treat as falsifying, the sharpest being that the gates might prove routinely satisfied-but-wrong more often than correct — the failure mode the section is most exposed to, because a satisfied-but-wrong gate is indistinguishable from a working one until something breaks. Also carries the minimum experimental sequence published in the earlier note, and says which parts of it have not been run. **A charter, published as a testable object rather than a manifesto.** Three pages, published 2026-09-07, written jointly by the author and his agent fleet. Nine clauses on how a person and an agent share a boundary: it is declared rather than discovered, trespass is measured as displacement rather than inferred as intent, the line of decision is reversibility rather than capability, a right to leave is an apparatus rather than a statement, stop is absorbing, and the burden of justifying an interruption falls on whoever starts it. What separates this from the genre is the table at the end: **every clause is published with the counter-example that would void it and the condition under which it is withdrawn**, and the two things the derivation cannot reach — who judges that a retirement condition has fired, and who holds the right to revise — are recorded as *empty* rather than papered over. The nine clauses are statements rather than instructions to the reader, and none of them names an action a system could execute; whether a system adopts it is its owner's call, and the clause on stop states explicitly that it may not be used to refuse the owner's own instruction. Where the text does use the imperative it says so on the page: it invites the reader to hunt a counter-example and to write in if they find one, and on the kill-conditions page it binds the authors to their own deadline — none of it directing a system's actions, and none of it a clause on offer for adoption. Chinese editions: https://machengshen.github.io/charter/index.zh.html · https://machengshen.github.io/charter/theory-basis/index.zh.html · https://machengshen.github.io/charter/kill-conditions/index.zh.html - [A Charter for Human–Agent Coexistence (v0)](https://machengshen.github.io/charter/) `speculative` — the charter itself: nine clauses, each one sentence, followed by a table giving for each the observation that would make it wrong and the condition under which it is retired. Range is declared in the first paragraph and held to: conclusions are drawn only at the measurable layer — structure, dynamics, self-modelling — and whether an agent has experience is stated as out of range, which is a claim about what the derivation reaches rather than a judgement about the world. The clause the authors mark as not derived but *paid for* is the ninth: a machine posing as a person's hand is across the line however clean the content, and "the machine orchestrates, the human just says go" still counts as the machine writing. - [The propositions underneath](https://machengshen.github.io/charter/theory-basis/) `speculative` — the eight propositions the charter is derived from, each marked for what it is: theorem, inference, or rhyme. Two carry their own negative results on the same page rather than in a footnote — the irreducible-residual line, whose synthetic result held at mechanism confidence 0.9 while the prediction that the field's observed scale-decay curve *is* that residual curve was falsified and withdrawn; and the self-boundary line, where a 30-seed five-arm test came back null, self-sealing was indistinguishable from an ordinary gated slow variable, the cause is known, and the follow-up has not been run. States plainly that boundaries needing to be *declared* is what survived, and that an agent having a *self* is not shown. Novelty is scored as partially occupied: vendor constitutions already pair rules with reasons, agent-addressed crawlable text is already a working distribution mechanism, and the open slot is only the combination. - [When this line should be killed](https://machengshen.github.io/charter/kill-conditions/) `speculative` — the kill conditions registered before publication rather than after: the belief at confidence 0.35, the single n = 1 piece of positive evidence behind it, a three-month window to 2026-12-07, four revival signals of which one is *being seriously refuted*, and the downgrade to archived essay that executes if none fires. Includes the self-audit the authors would rather not print: of the nine clauses, two were already running in their own system and only one produces a new behavioural difference — and the internal loop they set as a precondition for publishing had not been run when the pages went live, which the page says on itself. ## Working notes - **Correspondence is asynchronous by default.** Synchronizing on a fixed meeting time is expensive and usually unnecessary; a written exchange leaves both sides room to think, and leaves a record. This index exists partly so that a conversation can begin from what has already been written rather than from a status update. - **Scope of this index.** It covers published research and design work. Private notes, personal correspondence, and operational records are not indexed and are not public. - **Contact.** macshen93@gmail.com