Living Evidence Compendium

The Universal & Atomic Elements of Organization

A falsifiable scientific theory proposing that the same irreducible structures organize both thought and reality — from the quantum to the social scale.

What if the same structures that organize your thoughts also organize the world? Explore the hypothesis

O-Theory proposes that four irreducible structures — Distinctions, Systems, Relationships and Perspectives (DSRP) — form the universal grammar of organization. Rather than being merely useful ways of thinking, these structures are hypothesized to be the atomic elements from which every organized phenomenon emerges, whether in cognition, biology, physics, mathematics, society, or the cosmos.

DS RP
DistinctionsSystemsRelationshipsPerspectives
Identity ↔ OtherPart ↔ Whole Action ↔ ReactionPoint ↔ View
D := (i ↔ o)S := (p ↔ w) R := (a ↔ r)P := (ṗ ↔ v)

This is a living scientific evidence compendium: an open, continually evolving collection of independent empirical research, formal theory, mathematical proofs, cross-disciplinary analyses, applications, critiques, and proposed falsifications. Every entry is included because it supports, refines, challenges, or attempts to falsify the theory.

Scientific theories are strengthened not only by evidence that confirms their predictions, but also by surviving attempts to falsify them. This compendium brings both together: independent evidence from researchers who were not testing DSRP and proposed counterexamples evaluated against the formal theory.

One counterexample is enough to falsify O-Theory. Until then, the question remains: do the same four structures organize everything from quantum systems to human thought?

Try a demonstration yourself Why is this convergent evidence compelling? See why independent convergence is one of science’s strongest forms of evidence.

independent opportunities for the theory to fail.

Independent convergence is one of the strongest forms of scientific evidence because researchers arrive at the same conclusion while investigating different questions for different reasons.

The number is not the point. Researchers in different fields, studying different questions with different methods, repeatedly arrived at the same structural predictions—almost always without testing DSRP or using its language.

Discipline
Field

No studies match your search

Try a different term or clear the filters.

About this collection

This compendium began as the peer-reviewed literature review “A Literature Review of the Universal and Atomic Elements of Complex Cognition,” published in the Journal of Systems Thinking with 109 studies. That paper is the peer-reviewed foundation. What you see here is its living, continuously updated version . New studies are checked before they are added, and the collection now holds and keeps growing. Open any card to see what the researchers found, why it bears on DSRP, and where the original review discusses it, the fuller account.

Cabrera, D., Cabrera, L., & Cabrera, E. A Literature Review of the Universal and Atomic Elements of Complex Cognition. Journal of Systems Thinking. · Cornell University & Cabrera Research Lab.

How to cite this collection

APA
BibTeX

The collection is updated continuously, so the citation carries the date you consulted it rather than a study count — the count changes weekly, and putting it in the reference would make the same collection look like a different work to everyone who cites it. To cite a single claim or study, use its own address: every one has a permanent link.

About this record

This is the adversarial half of the compendium. Where the evidence track asks what converges on DSRP, this one asks what would end it: a single organized phenomenon whose structure needs a fifth pattern, a ninth element, or a fifth structural dynamic. It holds written up from candidates across territories, and resolutions — the general answers those cases settle against. Every case is published whether it held or failed, including the ones still open.

How to cite the counterexamples and resolutions

APA
BibTeX

Cite this rather than the evidence collection when the point is what survived attack. The two are separate works with separate addresses: one asks what converges on the theory, the other asks what would end it, and a reference to the first does not support a claim about the second. Case and resolution numbers change when the record is revised, so cite a case by its own permanent link rather than by number.

Know of work that belongs here?

This collection is meant to keep growing. Send us a study, paper, book or critique that bears on DSRP, whether it supports the theory or cuts against it, and we will read it and decide whether it belongs.

 

DSRP Evidence

Can mental fitness and its mechanisms be measured?

The research claim

Measuring thinking is where most frameworks stop. A systematic review of the mental fitness literature found six distinct measurement modes with no dominant standard, and a separate review of systems thinking frameworks found proliferation without empirical differentiation. Most accounts of good thinking cannot say how much of it someone has. The claim is that it can be measured, and that the thing being measured is a skill rather than a trait. That distinction decides everything downstream. A psychometric test is built to be stable — improvement would be error. An edumetric test is built to move, because the thing it measures is supposed to change. What is measured directly is the cognitive term: how well a person organizes information, and how accurately they judge their own organizing. The other three domains of fitness are specified as a measurement interface rather than a battery, with criteria for what counts as an admissible indicator. Three properties have to hold. The measure must be stable across occasions. It must relate to what it claims to measure rather than to something already measured by other means. And it must move when the person does — which for an edumetric instrument is the whole point, and the property with the least evidence behind it.

What would confirm it

Mental fitness and the moves that produce it can be measured — stably, in a way that relates to what is being claimed, and sensitively enough to register change. Read in symbols: reliability meets a stated threshold (τᵣ), validity meets a stated threshold (τᵥ), and sensitivity to change is greater than zero. All three are joined by ∧ because all three are required. An instrument that is stable and unrelated to anything is useless; so is one that relates to everything and never moves.

What would refute it

The measure is unstable, or measures something other than what it claims, or does not move when the person improves. Read in symbols: reliability below threshold, or (∨) validity below threshold, or sensitivity to change equal to zero. Any one of the three is fatal, which is why they are joined by or. The third is the most commonly fatal in practice: instruments that are stable and valid and register nothing when someone actually learns.

The evidence, and why this status

Two instruments, ten years apart, and the honest summary is that structure and reliability hold while sensitivity and independence do not. The current instrument shows high internal consistency and excellent model fit, with most reliable variance loading on a general factor and smaller pattern-specific subskills alongside it. That shape is itself a theoretical result rather than a psychometric convenience: the prediction was that the four patterns operate together rather than separately, so one dominant factor with real but secondary specialization is what should appear. A different shape would have been a problem for the theory, not just for the instrument. The weak points are reported rather than buried. Subscale reliabilities in the earlier and larger validation are modest, one pattern has never scored well in any version, and one fit index sits above its conventional threshold. The papers state plainly that the instrument has not been validated against IQ, aptitude tests, or existing metacognition and critical thinking measures — so incremental validity is claimed nowhere. One convergence is worth separating out because neither study was designed to produce it. The instrument finds confidence exceeding skill in every domain, by the widest margin for perspective. Separately, and by a completely different method, perspective is among the least-used patterns when people act with no instruction at all. People are worst at the move they are most confident about, found twice, by designs that fail in different ways. Underneath sits the item-level work: eighteen experiments establishing that each element can be elicited and scored on its own, which is what the instrument counts.

Supporting evidence (11 publications)

All research questions