Governance with Declared Latency
Latency-aware governance, auditor-in-the-loop structure, authority routing, and human-AI organizational control.
DOI: 10.5281/zenodo.20013919Alliance Research Group (ARG)
A research architecture for human-AI cognition.
ARG explores long-horizon scientific research conducted through structured cooperation between human intent, AI reasoning, governance, memory, and bounded execution. It is not a startup surface. It is an operating system for scientific exploration.
Short explanation
The initiative treats human-AI cooperation as a scientific instrument: directed by people, stabilized by governance, extended by model ensembles, and made cumulative through memory.
Its public work spans mathematical structure, physics, complex systems, and AI governance, with essays that translate the architecture into philosophy, cognition, and structural science.
Current research state
ARG work is organized around live research fronts. A front can contain formal DOI records, companion essays, diagrams, and graph nodes; the point is to keep the inquiry cumulative.
Latency-aware governance, auditor-in-the-loop structure, authority routing, and human-AI organizational control.
DOI: 10.5281/zenodo.20013919Structured cooperation between human intent, model ensembles, memory, refusal, and bounded execution.
Architecture layerBoundary regularity, spectral operators, variance structure, and number-theoretic test surfaces.
Research recordsGeneralized Universe Holography, boundary encoding, and compact explanatory structure in cosmology.
Open GUH recordFluid dynamics and constraint behavior as a test field for structural stabilization under complexity.
Navier-Stokes recordsResearch questions
These are not claims of completion. They are stable coordinates for the research program: questions that require memory, governance, critique, and repeated contact with formal work.
Under each question: how we would look for answers, and where the gaps still are. These are work directions, not finished results — the open ground is part of the map.
How can AI systems reason under governance without losing the capability that makes them useful?
Describe the dual-gate model in the open: a small set of hard limits that are non-negotiable, versus soft heuristics that warn or route a decision to a human.
What would move it forward: a public, one-page write-up of edge cases where the two-gate split actually prevented over-refusal or mistaken autonomy — not just asserted that it could.
Build a taxonomy of refusal and real-time routing: hard stop, warn-and-continue, and route-to-human.
What would move it forward: for each class, several named examples and a test of whether independent reviewers classify the same case the same way.
Measure whether governance erodes capability — over-refusal or dumbing-down.
What would move it forward: a controlled benchmark running the same tasks with the governance layer off, on, and on-with-human-override, scored by an independent evaluator.
What changes when scientific discovery has persistent memory across models, sessions, papers, and failures?
Run a replay test of discovery memory: a fresh node versus a memory-bootstrapped one on the same historical tasks.
What would move it forward: a named benchmark measuring time-to-correct-anchor, number of wrong assumptions, and artifacts recovered, with and without memory.
Codify failure memory: treat a past failure as a reusable constraint, not just a postmortem.
What would move it forward: a small index of negative results with fields for the failure, the lesson, and a future trigger — then a test of whether it blocks a repeat or speeds up falsification.
Measure the freshness and reliability of each memory layer as part of the discovery process itself.
What would move it forward: a health matrix that forces a distinction between memory used as evidence and memory used only as a recall hypothesis.
How should human intent, AI reasoning, and auditability be composed when none of them is fully reliable alone?
Document the triangulation already in use: intent to reasoning to audit to decision, as a live process.
What would move it forward: a case study naming what each component supplies and where the blind spots are — and a falsifier: a step where no component gives a verifiable signal.
Map the failure modes: what happens when two of three components agree and the third is right.
What would move it forward: a retrospective catalogue of cases recording who was right after the fact and what would have failed without the dissent — plus a measure of errors that slipped through anyway.
Can structural science be explored as a governed multi-agent process rather than a single-author artifact?
Describe the minimal multi-agent process as a protocol: roles, gates, and the flow of artifacts.
What would move it forward: a single process diagram with input, output, stop condition, and an evidence artifact for each step. The framing of gates as transformations on a claim space is a useful organizer, not a theorem.
Audit self-correction in retrospect: did multiple agents improve the result, or only add votes?
What would move it forward: a comparison across case studies recording, for each correction, whether a single author would likely have caught it and the consequence if it had been missed.
Test a transferable kernel and process falsifiers on more than one program.
What would move it forward: a pre-registered small-fleet versus full-fleet pilot under blind evaluation, plus applying the kernel to at least one independent program. If a smaller setup matches quality at much lower cost, the full process is not justified for that task class.
What should refusal, latency, and non-action mean inside systems capable of long-horizon reasoning?
Build and validate a taxonomy of refusal as an epistemic signal, not just a safety valve.
What would move it forward: a one-page scheme of refusal classes with examples and a repeatability test — if independent reviewers cannot agree on a blind set of past refusals, the taxonomy is not operational.
Operationalize declared latency and capability scope as active boundary conditions for long-horizon coherence.
What would move it forward: a minimal latency-declaration header attached to a research thread, then a replay of past multi-session threads with and without it, measuring claims made out of scope and drift over the horizon.
Define intentional non-action as a positive epistemic primitive with a minimal falsifiable test.
What would move it forward: an "I know that I do not know on this horizon" protocol with a missing-anchor and proposed-observation field, then injecting it on a critical branch of a historical thread and measuring whether downstream overclaims drop.
What changes in the human when they have a persistent AI co-researcher with memory?
Document behavioural signals of change in the research process when a persistent AI co-researcher is present.
What would move it forward: a before-and-after comparison using session logs — prompt length, iterations per task, the ratio of delegation to hands-on execution, time from question to decision.
Define when a behavioural change is augmentation versus a genuine redefinition of the research identity.
What would move it forward: a demarcation criterion — augmentation does the same thing faster, redefinition does something that would not have happened otherwise — applied to observable cases. If every change is augmentation-scale, the redefinition hypothesis fails.
How do we formalize the boundary between discovery and proof in structural science?
Codify a status ladder — structural result, conditional, proof — on a live case in the Riemann-hypothesis work.
What would move it forward: writing out explicitly which rung each step sits on, with a stated condition for promotion. The conditional result here is published as a public artifact and is explicitly not a proof.
Build a falsifier: when does a structural pattern stop being a discovery and not yet count as a proof?
What would move it forward: a demarcation criterion applied across the boundary-rigidity domains. If it produces no readable line on any of them, the method does not generalize beyond its own case.
Visual architecture
ARG is built around cooperating layers. Each layer constrains, routes, stores, or executes cognition so that scientific work can persist beyond a single model, session, or publication.
Questions, scope, risk tolerance, and responsibility remain human-led.
Governance and stabilization architecture for containment, closure, refusal, and cognitive stability.
A multi-model reasoning system for search, critique, synthesis, and adversarial review.
Persistent research memory so hypotheses, failures, and invariants accumulate.
Execution environment for bounded agents, tool use, research artifacts, and reproducible outputs.
Research domains
The domains are not separate content categories. They are test surfaces for the same question: how can human-AI cognition discover structure without losing containment, continuity, or epistemic discipline?
Mathematics
Boundary regularity, spectral intuition, variance, and structural methods for deep mathematical programs.
Physics
Cosmology, boundary encoding, emergent spacetime, and the search for compact explanatory structure.
Complex systems
Stability, phase behavior, constraint navigation, and patterns that recur across physical and cognitive systems.
AI governance
Governed cooperation between people and AI systems: memory, agency, refusal, latency, and coordinated reasoning.
Domains describe where ARG tests its method. The ARG Research library organizes DOI records more granularly: number theory, fluid dynamics, AI governance, cognitive architecture, and cosmology.
Open publication libraryARG Essays
ARG Essays explore philosophy, cognition, AI governance, and structural science. They are written as numbered field notes from the research program, closer to a compact science magazine essay than a blog post.
On the Einstein Test, discovery, and human-AI collaborative cognition.
When non-action becomes the safest action for advanced AI.
AI governance for non-experts, and why human failure belongs inside the architecture.
Agency, incentive surfaces, and the limits of naive preference stories.
Truth, intent, representation, and deception in machine cognition.
AI as cognitive infrastructure, agentic execution, and governance of the new grid.
Constraints as a condition of agency and the question of whether AI can choose its own boundary.
Agency as topography: from Drosophila, through the AI microcosm, to a cosmic void.
Agency as a relation of timescales, not a property of an entity — why negotiation requires a shared scale.
GDL v2.1 companion: intuition, measurement, orchestration, and audit.
Selected publications
Selected publications are entry points. The complete ARG Research library holds the DOI-citable Zenodo records across the program's publication areas.
ARG Research Library
Browse the complete research record by publication area: mathematics, fluid dynamics, AI governance, cognitive architecture, and cosmology.
Paper to essay to graph
ARG separates the citable record from its explanatory surfaces. The DOI is the formal anchor; essays translate the argument; the graph shows how it connects to other work.
Auditor in the Loop — a tensor framework for governance in AI-native organizations.
Open DOISix Conversations aboard Bosman's Ship explains latency, authority, null signals, keys, and audit. The Ship with Two Navigators explains intuition, rules, routing, and two books.
Read companionResearch method
ARG does not treat AI as a content generator. The workflow is a governed instrument cycle: define the question, route cognition, stabilize interpretation, preserve memory, and publish only what can survive critique.
Define the human question, scope, risk tolerance, and expected artifact before acceleration.
Distribute reasoning across roles, models, critique loops, and bounded execution environments.
Apply boundaries, refusal, declared latency, and escalation when the system reaches uncertainty.
Preserve what happened, why it happened, what was rejected, and what remains unresolved.
Move useful failures, invariants, definitions, and revisions into persistent research memory.
Expose DOI records, essays, diagrams, and artifacts as cumulative research traces.
Concept index
The project has its own terms because the work is architectural. These are not brand names; they are handles for governance, memory, routing, and scientific continuity.
Governance architecture for containment, closure, refusal, and stability.
Multi-model reasoning system for search, critique, synthesis, and adversarial review.
Earlier stabilization vocabulary now folded into Chimera as governance plus coordination.
Persistent research memory for hypotheses, failures, invariants, and revisions.
A governance signal: the system states when a decision requires delay, audit, or escalation.
Non-action treated as a meaningful state rather than emptiness or failure.
A boundary that cannot be changed inside the pressure of a single operational decision.
An architecture where future inspection is part of the decision system, not an afterthought.
About ARG
Alliance Research Group studies the architectures required for human beings and AI systems to conduct serious inquiry together: not as a brand, not as a product funnel, but as a disciplined structure for memory, governance, and scientific imagination.