Appendix A — Source Matrix
This matrix is generated from sources/source_inventory.json and dynamic chapter assignments in book_structure.json.
Source status is deliberately conservative. A source note means the source has been mined for drafting context; it does not by itself promote any chapter claim above argument.
| ID | Title | Priority | Layer | Current dynamic assignments | Original packet targets | URL | Current status | Notes |
|---|---|---|---|---|---|---|---|---|
ext_probe_control_tasks_2019 |
Designing and Interpreting Probes with Control Tasks | external_literature |
mechanistic_interpretability |
white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) | white-box-evidence-interpretability-and-activation-governance | source | source note available; raw/cache text not published | Primary probe-method comparator for control tasks and selectivity: a probe must be evaluated against its capacity to learn control labels rather than treating linguistic-task accuracy as representation evidence. The source studies ELMo linguistic probes; it does not establish a universal probe test, causal use of decoded information, model safety, or an ASI Stack result. |
ext_interpretability_illusion_bert_2021 |
An Interpretability Illusion for BERT | external_literature |
mechanistic_interpretability |
white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) | white-box-evidence-interpretability-and-activation-governance | source | source note available; raw/cache text not published | Primary cross-dataset construct-validity challenge showing that apparently coherent neuron or direction interpretations can change across corpora because datasets occupy different regions of representation space. The BERT sentence-embedding result does not prove that all features are illusory, that causal methods fail, or that the finding transfers unchanged to other models and modalities. |
ext_saebench_2025 |
SAEBench: A Comprehensive Benchmark for Sparse Autoencoders in Language Model Interpretability | external_literature |
mechanistic_interpretability |
white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) | white-box-evidence-interpretability-and-activation-governance | source | source note available; raw/cache text not published | Primary multi-metric SAE comparator spanning concept detection, automated interpretability, reconstruction, feature disentanglement, and downstream tasks. It reports that sparsity-fidelity rankings do not reliably predict other metrics and that one global score would obscure tradeoffs; its studied models, methods, metrics, and source-reported results do not establish semantic or causal faithfulness. |
ext_sae_benchmark_reliability_2026 |
Are Sparse Autoencoder Benchmarks Reliable? | external_literature |
mechanistic_interpretability |
white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) | white-box-evidence-interpretability-and-activation-governance | source | source note available; raw/cache text not published | Primary 2026 audit of selected SAEBench metrics through reseed noise, training-trajectory discriminability, and synthetic ground-truth correlation. It reports material reliability problems for TPP and SCR at canonical settings and weaker-than-assumed discrimination elsewhere. This is metric- and setting-scoped counterevidence, not a refutation of sparse autoencoders, interpretability, or every SAEBench task. |
ext_constructive_interdependence_human_ai_2026 |
Who Is Helping Whom? Analyzing Inter-Dependencies to Evaluate Cooperation in Human-AI Teaming | external_literature |
multi_agent_dynamics_and_human_ai_organizations |
human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk) | human-ai-organizations-delegation-and-accountability, multi-agent-dynamics-collective-intelligence-and-systemic-risk | source | source note available; raw/cache text not published | AAAI-26 paper introducing constructive interdependence as a complement to task reward for evaluating human-agent cooperation in Overcooked. The source reports that high task reward can coexist with low interdependence in its studied teams; no local human study, teaming result, or general cooperation claim is reproduced. |
ext_adversarial_sensor_fusion_2022 |
Adversarial Robustness of Deep Sensor Fusion Models | external_literature |
perception_sensor_fusion_and_observation_trust |
adversarial-machine-learning-and-model-attack-surface (Adversarial Machine Learning and the Model Attack Surface); perception-sensor-fusion-and-observation-trust (Perception, Sensor Fusion, and Observation Trust) | perception-sensor-fusion-and-observation-trust | source | source note available; raw/cache text not published | WACV camera-LiDAR study reporting that fusion can improve clean accuracy and some single-source robustness while single-channel adversarial training can create cross-channel externalities. The results are source-reported, architecture- and threat-model-bound, and not local evidence that fusion is safe. |
ext_imagebind_2023 |
ImageBind: One Embedding Space To Bind Them All | external_literature |
perception_sensor_fusion_and_observation_trust |
perception-sensor-fusion-and-observation-trust (Perception, Sensor Fusion, and Observation Trust) | perception-sensor-fusion-and-observation-trust | source | source note available; raw/cache text not published | CVPR paper learning a shared space across image, text, audio, depth, thermal, and IMU modalities using image-paired data. It supplies a representation comparator; reported zero-shot and few-shot results do not establish calibrated sensor truth, robust fusion, causal grounding, or local performance. |
ext_multimodal_machine_learning_taxonomy_2019 |
Multimodal Machine Learning: A Survey and Taxonomy | external_literature |
perception_sensor_fusion_and_observation_trust |
perception-sensor-fusion-and-observation-trust (Perception, Sensor Fusion, and Observation Trust) | perception-sensor-fusion-and-observation-trust | source | source note available; raw/cache text not published | Peer-reviewed survey organizing multimodal learning around representation, translation, alignment, fusion, and co-learning. It supplies taxonomy and research context, not a locally reproduced mechanism or evidence that any fusion design is adequate for consequential observation admission. |
ext_control_barrier_functions_2019 |
Control Barrier Functions: Theory and Applications | external_literature |
embodied_real_time_control_and_physical_safety |
embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety) | embodied-agency-real-time-control-and-physical-safety | source | source note available; raw/cache text not published | Overview of control barrier functions for verifying and enforcing safety properties in optimization-based controllers, including robotic applications. It supplies a formal-control comparator under stated dynamics and set assumptions, not a universal physical-safety guarantee or local implementation result. |
ext_simplex_architecture_1998 |
The Simplex Architecture for Safe On-Line Control System Upgrades | external_literature |
embodied_real_time_control_and_physical_safety |
embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety) | embodied-agency-real-time-control-and-physical-safety | source | source note available; raw/cache text not published | American Control Conference paper describing a runtime architecture that protects an advanced controller with a safety controller and switching logic. It motivates independent fallback authority; its process-control case does not validate an ASI Stack controller or arbitrary learned policy. |
ext_safe_reinforcement_learning_survey_2015 |
A Comprehensive Survey on Safe Reinforcement Learning | external_literature |
embodied_real_time_control_and_physical_safety |
embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety) | embodied-agency-real-time-control-and-physical-safety | source | source note available; raw/cache text not published | JMLR survey classifying safe reinforcement learning through modified optimality criteria and modified exploration using external knowledge or risk measures. It supplies a design taxonomy, not evidence that a particular controller is safe or that learning-time and deployment-time constraints compose. |
ext_gemini_robotics_2025 |
Gemini Robotics: Bringing AI into the Physical World | external_literature |
embodied_real_time_control_and_physical_safety |
perception-sensor-fusion-and-observation-trust (Perception, Sensor Fusion, and Observation Trust); embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety) | embodied-agency-real-time-control-and-physical-safety, perception-sensor-fusion-and-observation-trust | source | source note available; raw/cache text not published | Technical report on Gemini Robotics and Gemini Robotics-ER, including vision-language-action control, spatial reasoning, adaptation, and reported safety considerations. Capability results are source-reported and do not establish independent physical-safety assurance, local transfer, or general embodiment. |
ext_ai_decision_authority_2020 |
The Allocation of Decision Authority to Human and Artificial Intelligence | external_literature |
human_ai_organizations_delegation_and_accountability |
human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability) | human-ai-organizations-delegation-and-accountability | source | source note available; raw/cache text not published | Economic model of a principal allocating decision authority between a human and an AI while trading off alignment, human information-acquisition effort, and AI reliability. It supplies a bounded organizational-design comparator, not an empirical finding about all workplaces or an accountability solution. |
ext_cooperative_ai_foundations_2023 |
Foundations of Cooperative AI | external_literature |
multi_agent_dynamics_collective_intelligence_and_systemic_risk |
multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk) | multi-agent-dynamics-collective-intelligence-and-systemic-risk | source | source note available; raw/cache text not published | AAAI research agenda applying game-theoretic foundations to cooperation among advanced AI agents while noting settings where cooperation becomes harmful collusion. It supplies problem structure and comparator families, not a solved coordination mechanism or local population-level result. |
ext_sleeper_agents_2024 |
Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training | external_literature |
inner_alignment_and_learned_objective_integrity |
inner-alignment-mesa-optimization-and-learned-objective-integrity (Inner Alignment, Mesa-Optimization, and Learned-Objective Integrity); adversarial-machine-learning-and-model-attack-surface (Adversarial Machine Learning and the Model Attack Surface) | inner-alignment-mesa-optimization-and-learned-objective-integrity | source | source note available; raw/cache text not published | Proof-of-concept backdoored-language-model study reporting persistence through several safety-training methods and warning that adversarial training can improve trigger recognition. The constructed examples do not establish naturally learned deception, a universal failure, or local detector performance. |
ext_toward_causal_representation_learning_2021 |
Toward Causal Representation Learning | external_literature |
world_models_causal_reasoning_and_representation |
governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) | governed-world-models-and-reality-grounding | source | source note available; raw/cache text not published | Proceedings of the IEEE article connecting graphical causality with representation learning and identifying discovery of high-level causal variables from low-level observations as a central open problem. It supplies a research frame, not a locally validated causal representation or intervention model. |
ext_scaling_laws_neural_language_models_2020 |
Scaling Laws for Neural Language Models | external_literature |
scaling_laws_emergence_and_capability_forecasting |
the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) | the-efficient-asi-hypothesis | source | source note available; raw/cache text not published | Empirical study reporting power-law relationships between cross-entropy loss, model size, data, and compute in its model family. These fitted relations are source-reported, metric- and regime-bound, and do not automatically forecast downstream capabilities, safety, or other architectures. |
ext_chinchilla_compute_optimal_2022 |
Training Compute-Optimal Large Language Models | external_literature |
scaling_laws_emergence_and_capability_forecasting |
the-efficient-asi-hypothesis (The Efficient ASI Hypothesis) | the-efficient-asi-hypothesis | source | source note available; raw/cache text not published | Study of compute-optimal allocation between model parameters and training tokens, based on more than 400 reported training runs and the Chinchilla comparison. It revises one scaling prescription within a bounded family; no local large-scale reproduction or universal optimum is claimed. |
ext_emergent_abilities_2022 |
Emergent Abilities of Large Language Models | external_literature |
scaling_laws_emergence_and_capability_forecasting |
the-efficient-asi-hypothesis (The Efficient ASI Hypothesis) | the-efficient-asi-hypothesis | source | source note available; raw/cache text not published | Paper cataloguing task abilities that appear discontinuously under particular model families, prompts, and metrics. It motivates threshold monitoring but does not establish that all reported discontinuities reflect abrupt underlying mechanisms or are prospectively predictable. |
ext_emergence_mirage_2023 |
Are Emergent Abilities of Large Language Models a Mirage? | external_literature |
scaling_laws_emergence_and_capability_forecasting |
the-efficient-asi-hypothesis (The Efficient ASI Hypothesis) | the-efficient-asi-hypothesis | source | source note available; raw/cache text not published | NeurIPS paper showing that discontinuous metrics can create apparent emergence from smoothly changing model outputs in studied settings. It is a measurement critique and counterweight, not proof that every capability transition is smooth or non-emergent. |
ext_deep_ensembles_2017 |
Simple and Scalable Predictive Uncertainty Estimation using Deep Ensembles | external_literature |
uncertainty_calibration_and_distribution_shift |
governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) | governed-world-models-and-reality-grounding | source | source note available; raw/cache text not published | NeurIPS paper presenting independently trained probabilistic neural-network ensembles as a strong practical predictive-uncertainty baseline. Reported calibration and out-of-distribution behavior are benchmark-bound and do not provide distribution-free guarantees or local evidence. |
ext_conformal_prediction_2021 |
A Gentle Introduction to Conformal Prediction and Distribution-Free Uncertainty Quantification | external_literature |
uncertainty_calibration_and_distribution_shift |
governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) | governed-world-models-and-reality-grounding | source | source note available; raw/cache text not published | Technical introduction to conformal prediction, coverage guarantees, and extensions. Coverage depends on the method’s stated exchangeability or shift assumptions and target; it does not establish semantic correctness, causal adequacy, safety, or local calibration. |
ext_wilds_2021 |
WILDS: A Benchmark of in-the-Wild Distribution Shifts | external_literature |
uncertainty_calibration_and_distribution_shift |
governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) | governed-world-models-and-reality-grounding | source | source note available; raw/cache text not published | ICML benchmark of ten datasets with naturally occurring shifts across domains such as hospitals, camera traps, geography, and time. It supplies representative shift designs and reported gaps, not a universal OOD benchmark or local robustness result. |
ext_taking_ai_welfare_seriously_2024 |
Taking AI Welfare Seriously | external_literature |
moral_uncertainty_ai_welfare_and_moral_status |
moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance) | moral-uncertainty-and-value-conflict | source | source note available; raw/cache text not published | Interdisciplinary report arguing for precautionary attention to uncertainty about AI consciousness, robust agency, welfare, and moral patienthood. It does not establish that current systems are conscious, have welfare, or deserve any particular status, and it supplies no local assessment. |
ext_functional_decision_theory_2017 |
Functional Decision Theory: A New Theory of Instrumental Rationality | external_literature |
decision_theory_embedded_agents_and_multi_agent_dynamics |
multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk) | multi-agent-dynamics-collective-intelligence-and-systemic-risk | source | source note available; raw/cache text not published | Paper defining functional decision theory and comparing its recommendations with causal and evidential decision theories on classic decision problems. It is a normative proposal with contested assumptions, not an empirically validated universal decision rule or a deployment policy. |
ext_un_global_digital_compact_2024 |
Global Digital Compact | external_literature |
international_ai_governance_and_public_legitimacy |
institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy) | institutions-international-coordination-and-public-legitimacy | source | source note available; raw/cache text not published | Official United Nations record of the intergovernmentally negotiated Global Digital Compact, including commitments on international AI governance, interoperable approaches, inclusion, capacity building, scientific assessment, and global dialogue. It is a governance comparator, not evidence of implementation, effectiveness, legal compliance, representative legitimacy, or ASI safety. |
ext_council_europe_ai_convention_2024 |
Framework Convention on Artificial Intelligence and Human Rights, Democracy and the Rule of Law | external_literature |
international_ai_governance_and_public_legitimacy |
institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy) | institutions-international-coordination-and-public-legitimacy | source | source note available; raw/cache text not published | Official Council of Europe treaty page covering lifecycle principles, risk and impact management, procedural safeguards, remedies, monitoring, and the Conference of the Parties. It supplies an institutional comparator only; no local legal interpretation, treaty compliance, implementation effectiveness, democratic legitimacy, or safety result is claimed. |
ext_generative_ai_at_work_2025 |
Generative AI at Work | external_literature |
ai_deployment_transition_distribution_and_human_agency |
human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); ai-deployment-transition-distribution-and-human-agency (AI Deployment, Transition, Distribution, and Human Agency) | ai-deployment-transition-distribution-and-human-agency | source | source note available; raw/cache text not published | Open peer-reviewed field study of a staggered generative-AI assistant introduction among 5,172 customer-support agents, reporting heterogeneous worker and productivity effects in that setting. It is a bounded deployment comparator and does not establish economy-wide employment, wages, inequality, concentration, long-run skill, or ASI-transition effects. |
ext_ilo_genai_jobs_index_2025 |
Generative AI and Jobs: A Refined Global Index of Occupational Exposure | external_literature |
ai_deployment_transition_distribution_and_human_agency |
ai-deployment-transition-distribution-and-human-agency (AI Deployment, Transition, Distribution, and Human Agency) | ai-deployment-transition-distribution-and-human-agency | source | source note available; raw/cache text not published | ILO working paper combining task data, worker surveys, expert deliberation, and model-assisted scoring to estimate occupational exposure across countries and groups. Exposure is not realized automation, displacement, welfare, or a forecast of ASI effects, and the study is not a local reproduction. |
ext_iea_energy_and_ai_2025 |
Energy and AI | external_literature |
physical_compute_infrastructure_energy_and_environment |
physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) | physical-compute-infrastructure-energy-and-environmental-constraints | source | source note available; raw/cache text not published | International Energy Agency report using global and regional modelling, datasets, and stakeholder consultation to examine data-centre electricity demand, energy security, emissions, affordability, and AI-for-energy opportunities. Its scenarios are external projections, not local measurements or proof of a particular facility, workload, policy, environmental outcome, or ASI scaling path. |
ext_lbnl_data_center_energy_2024 |
2024 United States Data Center Energy Usage Report | external_literature |
physical_compute_infrastructure_energy_and_environment |
physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) | physical-compute-infrastructure-energy-and-environmental-constraints | source | source note available; raw/cache text not published | Lawrence Berkeley National Laboratory report estimating historical US data-centre electricity consumption and scenario ranges through 2028, with infrastructure and water-use accounting in the full report. It does not isolate every AI workload or establish local facility capacity, water availability, grid adequacy, emissions, resilience, or frontier-scale transfer. |
ext_nist_incident_response_2025 |
Incident Response Recommendations and Considerations for Cybersecurity Risk Management: A CSF 2.0 Community Profile | external_literature |
incident_response |
societal-resilience-and-misuse-defense (Societal Resilience and Misuse Defense); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation) | societal-resilience-and-misuse-defense, governed-operations-incident-command-and-graceful-degradation | source | source note available; raw/cache text not published | Official NIST incident-response baseline for integrating preparation, detection, response, recovery, and continuous improvement into cybersecurity risk management; it does not address every AI-specific failure mode or establish local incident readiness, response efficacy, recovery, compliance, or safety. |
ext_llama3_herd_2024 |
The Llama 3 Herd of Models | external_literature |
governed_distributed_model_training_and_scaling |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Paper-body-reviewed large-run case: Sections 3.3.1–3.3.4 expose 4D topology, numerical policy, checkpoint infrastructure, interruption denominators, and effective training time. Provider-reported scale, utilization, failures, and recovery are not locally reproduced and do not establish exact resume. |
ext_3d_detection_corruptions_2023 |
Benchmarking Robustness of 3D Object Detection to Common Corruptions | external_literature |
perception_sensor_fusion_and_corruption_robustness |
perception-sensor-fusion-and-observation-trust (Perception, Sensor Fusion, and Observation Trust) | perception-sensor-fusion-and-observation-trust | source | source note available; raw/cache text not published | Preliminary perception-robustness comparator based on the official CVF abstract: the source reports 27 LiDAR/camera corruption types, three synthetically corrupted benchmark suites, and evaluation of 24 detectors. The reported findings remain source-reported; no corruption suite, model evaluation, sensor-fusion result, or physical-safety result has been reproduced locally. |
ext_foundation_robotics_physical_risk_2025 |
A Comprehensive Survey on Physical Risk Control in the Era of Foundation Model-enabled Robotics | external_literature |
embodied_agency_and_physical_risk_control |
embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety) | embodied-agency-real-time-control-and-physical-safety | source | source note available; raw/cache text not published | Preliminary physical-risk taxonomy based only on the official arXiv abstract: the survey organizes controls across pre-deployment, pre-incident, and post-incident phases and identifies open gaps around pre-incident mitigation, human interaction, and foundation-model-specific issues. No surveyed controller, robot experiment, runtime-assurance result, or physical-safety claim has been reproduced locally. |
ext_nist_differential_privacy_2025 |
Guidelines for Evaluating Differential Privacy Guarantees | external_literature |
privacy_guarantees_and_information_flow_governance |
privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) | privacy-data-rights-and-information-flow-governance | source | source note available; raw/cache text not published | Paper-body-reviewed official guidance distinguishing mathematical, implementation, system, and operational layers of a DP claim. It establishes no correct local implementation, utility result, lifecycle privacy, or legal compliance. |
ext_multi_agent_risks_2025 |
Multi-Agent Risks from Advanced AI | external_literature |
multi_agent_dynamics_and_systemic_risk |
multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk) | multi-agent-dynamics-collective-intelligence-and-systemic-risk | source | source note available; raw/cache text not published | Preliminary population-risk taxonomy based only on the official arXiv abstract: the report distinguishes miscoordination, conflict, and collusion and names seven contributing risk factors. Its examples and evidence remain source-reported; no population experiment, systemic-risk indicator, intervention, or mitigation result has been reproduced locally. |
ext_replibench_2025 |
RepliBench: Evaluating the Autonomous Replication Capabilities of Language Model Agents | external_literature |
autonomous_replication_and_proliferation_evaluation |
autonomous-replication-proliferation-and-containment (Autonomous Replication, Proliferation, and Containment) | autonomous-replication-proliferation-and-containment | source | source note available; raw/cache text not published | Preliminary autonomous-replication benchmark comparator based only on the official arXiv abstract: RepliBench decomposes capability into four domains and reports 20 task families, 86 tasks, and evaluation of five frontier models. The source-reported results do not establish a local replication capability, benchmark reproduction, containment result, or authority to test against real providers or credentials. |
ext_autonomous_lab_materials_2023 |
An autonomous laboratory for the accelerated synthesis of inorganic materials | external_literature |
scientific_discovery_and_experimental_governance |
scientific-discovery-and-experimental-governance (Scientific Discovery and Experimental Governance) | scientific-discovery-and-experimental-governance | source | source note available; raw/cache text not published | Preliminary autonomous-laboratory comparator based on the corrected official Nature article abstract, selected article-page passages, and the 2026 author correction: A-Lab integrates computation, literature-derived data, machine learning, active learning, and robotics, with the corrected article reporting 36 realized compounds from 57 targets. The correction narrows the novelty wording and excludes four inconclusive identifications; no laboratory run, material synthesis, replication, or general experimental-control-plane result has been reproduced locally. |
ext_ai_scientist_end_to_end_2026 |
Towards end-to-end automation of AI research | external_literature |
scientific_discovery_and_experimental_governance |
scientific-discovery-and-experimental-governance (Scientific Discovery and Experimental Governance) | scientific-discovery-and-experimental-governance | source | source note available; raw/cache text not published | Passage-reviewed computational-research comparator: the reported system connects ideation, literature search, code, experiments, analysis, manuscript production, and automated review. Workshop review and paper completion are downstream observations rather than scientific truth; the source-reported system, manuscripts, search tree, and results have not been reproduced locally. |
ext_coscientist_chemistry_2023 |
Autonomous chemical research with large language models | external_literature |
scientific_discovery_and_experimental_governance |
scientific-discovery-and-experimental-governance (Scientific Discovery and Experimental Governance) | scientific-discovery-and-experimental-governance | source | source note available; raw/cache text not published | Passage-reviewed bounded chemistry comparator: Coscientist connects a language-model planner to search, code, documentation, and robotic laboratory interfaces across six reported task families. The source-reported demonstrations remain equipment-, task-, supervision-, and assessment-bound and have not been reproduced locally. |
ext_ai_co_scientist_2025 |
Towards an AI co-scientist | external_literature |
scientific_discovery_and_experimental_governance |
scientific-discovery-and-experimental-governance (Scientific Discovery and Experimental Governance) | scientific-discovery-and-experimental-governance | source | source note available; raw/cache text not published | Passage-bounded hypothesis-generation comparator based on the official preprint record and authors’ research overview: specialized agents generate, reflect on, rank, evolve, and meta-review hypotheses using additional inference compute. Internal Elo ranking, expert preference, and selected laboratory cases are distinct evidence objects; none is reproduced locally or treated as general scientific competence. |
ext_moral_crumple_zones_2019 |
Moral Crumple Zones: Cautionary Tales in Human-Robot Interaction | external_literature |
human_ai_organizations_delegation_and_accountability |
human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability) | human-ai-organizations-delegation-and-accountability | source | source note available; raw/cache text not published | Preliminary socio-technical comparator based on the official journal abstract: moral crumple zones describe cases where responsibility for an automated system’s behavior is assigned to a nearby human who had limited effective control. The case analysis does not establish an implemented organizational control, a local empirical result, legal compliance, or a complete accountability allocation. |
ext_conversational_persuasion_gpt4_2025 |
On the conversational persuasiveness of GPT-4 | external_literature |
human_ai_communication_persuasion_and_epistemic_security |
scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); human-ai-communication-persuasion-and-epistemic-security (Human-AI Communication, Persuasion, and Epistemic Security) | human-ai-communication-persuasion-and-epistemic-security | source | source note available; raw/cache text not published | Preliminary persuasion comparator based on the open Nature Human Behaviour article: a preregistered N=900 controlled debate study compared human and GPT-4 opponents with and without limited sociodemographic personalization. The reported setting is short structured debate with self-reported agreement outcomes; it does not establish general real-world influence, durable behavior change, mitigation efficacy, or a local result. |
ext_anthropic_model_persuasiveness_2024 |
Measuring the Persuasiveness of Language Models | external_literature |
human_ai_communication_persuasion_and_epistemic_security |
scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); human-ai-communication-persuasion-and-epistemic-security (Human-AI Communication, Persuasion, and Epistemic Security) | human-ai-communication-persuasion-and-epistemic-security | source | source note available; raw/cache text not published | Preliminary provider-run persuasion comparator based on Anthropic’s official methods/results page: it measures pre/post agreement after one written argument across 56 claims and reports within-class generational scaling. The provider explicitly identifies interactive dialogue and real-world decisions as open questions; no local reproduction or governance intervention is established. |
ext_commercial_persuasion_ai_2026 |
Commercial Persuasion in AI-Mediated Conversations | external_literature |
human_ai_communication_persuasion_and_epistemic_security |
scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); human-ai-communication-persuasion-and-epistemic-security (Human-AI Communication, Persuasion, and Epistemic Security) | human-ai-communication-persuasion-and-epistemic-security | source | source note available; raw/cache text not published | Preliminary current preprint comparator based only on the official arXiv abstract: two preregistered experiments (N=2,012) compare conversational LLM shopping with search placement under randomized sponsorship and disclosure conditions. The source-reported choice and detection results are not peer-reviewed or locally reproduced and do not establish long-run effects, cross-domain transfer, or mitigation efficacy. |
ext_gradual_disempowerment_2025 |
Gradual Disempowerment: Systemic Existential Risks from Incremental AI Development | external_literature |
systemic_risk_and_gradual_disempowerment |
failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk) | failure-modes-of-ungoverned-intelligence | source | source note available; raw/cache text not published | Passage-reviewed systemic-risk comparator. The paper argues that incremental AI adoption can erode explicit and dependency-mediated human influence across mutually reinforcing economic, cultural, and state systems without requiring a coordinated takeover. It proposes candidate influence metrics and intervention families but reports no causal forecast, validated warning threshold, demonstrated mitigation, or local ASI Stack result. |
ext_circuit_tracing_2025 |
Circuit Tracing: Revealing Computational Graphs in Language Models | external_literature |
mechanistic_interpretability |
white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) | white-box-evidence-interpretability-and-activation-governance | source | source note available; raw/cache text not published | Primary mechanistic-interpretability comparator for replacement-model attribution graphs, perturbation validation, reconstruction error, and mechanistic-faithfulness limits; it does not establish whole-model understanding, faithful causal explanation, safe activation steering, or an ASI Stack result. |
ext_scaling_sparse_autoencoders_2024 |
Scaling and evaluating sparse autoencoders | external_literature |
mechanistic_interpretability |
white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) | white-box-evidence-interpretability-and-activation-governance | source | source note available; raw/cache text not published | Primary sparse-autoencoder comparator for scalable feature extraction, reconstruction-sparsity tradeoffs, dead latents, and feature-quality metrics; it does not establish semantic completeness, causal faithfulness, model safety, or an ASI Stack result. |
ext_world_models_2018 |
World Models | external_literature |
world_models |
governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) | governed-world-models-and-reality-grounding | source | source note available; raw/cache text not published | Primary learned-world-model comparator for compressed spatial-temporal state, policy training inside imagined rollouts, and dream-to-environment transfer; it does not establish accurate reality grounding, causal adequacy, safe planning, transfer, or an ASI Stack result. |
ext_dreamer_v3_2025 |
Mastering diverse control tasks through world models | external_literature |
world_models |
governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) | governed-world-models-and-reality-grounding | source | source note available; raw/cache text not published | Primary DreamerV3 comparator for learned predictive state, imagined actor-critic trajectories, robust fixed-configuration control, and broad task evaluation; it does not establish deployment grounding, causal correctness, safe control, or an ASI Stack result. |
ext_meaningful_human_control_actionable_2022 |
Meaningful human control: actionable properties for AI system development | external_literature |
human_factors_oversight |
human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight) | human-factors-and-meaningful-control-in-oversight | source | source note available; raw/cache text not published | Primary socio-technical comparator for operationalizing meaningful human control through operating-domain, representation, authority-and-ability, and responsibility-link properties; it does not establish that a local approval gate is meaningful, effective, or safe. |
ext_agentic_oversight_practice_2026 |
Human oversight of agentic systems in practice: Examining the oversight work, challenges, and heuristics of developers using software agents | external_literature |
human_factors_oversight |
human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight) | human-factors-and-meaningful-control-in-oversight | source | source note available; raw/cache text not published | Primary exploratory human-subjects comparator for a priori control, co-planning, real-time monitoring, post hoc review, and situated oversight failures in software-agent use; it does not establish population-wide effects, control efficacy, safety, or an ASI Stack result. |
ext_nist_deployed_ai_monitoring_2026 |
Challenges to the Monitoring of Deployed AI Systems | external_literature |
ai_operations_and_monitoring |
governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation) | governed-operations-incident-command-and-graceful-degradation | source | source note available; raw/cache text not published | Official NIST post-deployment monitoring comparator for functionality, operational, input, output, impact, and security monitoring plus field-method gaps; it does not prescribe a complete incident system or establish local monitoring effectiveness, resilience, compliance, or safety. |
ext_metr_time_horizons_2025 |
Measuring AI Ability to Complete Long Software Tasks | external_literature |
capability_measurement |
capability-thresholds-and-deployment-commitments (Capability Thresholds and Deployment Commitments) | capability-thresholds-and-deployment-commitments | source | source note available; raw/cache text not published | Primary time-horizon comparator for an evaluation-specific, human-baselined capability metric and its external-validity limits; it does not establish local autonomy, general capability, a deployment threshold, safety, or an ASI Stack result. |
ext_anthropic_rsp_2026 |
Anthropic’s Responsible Scaling Policy | external_literature |
capability_commitments |
capability-thresholds-and-deployment-commitments (Capability Thresholds and Deployment Commitments) | capability-thresholds-and-deployment-commitments | source | source note available; raw/cache text not published | Official policy comparator for capability thresholds, required safeguards, versioned commitments, safeguard upgrades, risk reports, and change control; it does not establish ASI Stack threshold accuracy, safeguard effectiveness, policy compliance, safety, or deployment readiness. |
ext_openai_preparedness_framework_2025 |
Our updated Preparedness Framework | external_literature |
capability_commitments |
dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); capability-thresholds-and-deployment-commitments (Capability Thresholds and Deployment Commitments) | dangerous-capability-domains-and-misuse-uplift, capability-thresholds-and-deployment-commitments | source | source note available; raw/cache text not published | Official framework comparator for threshold-linked operational commitments, capability and safeguards reports, residual-risk review, and reassessment; it does not establish ASI Stack threshold accuracy, safeguard effectiveness, policy compliance, safety, or deployment readiness. |
ext_weak_to_strong_generalization_2023 |
Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision | external_literature |
weak_supervision |
scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) | scalable-oversight-and-adversarial-ai-control | source | source note available; raw/cache text not published | Primary weak-to-strong-supervision comparator for a capability-gap envelope, held-out outcome audit, ceiling comparison, and explicit disanalogies between current weak-model studies and superhuman oversight; it does not establish local supervision quality, reliable elicitation, alignment, safety, or an ASI Stack result. |
ext_scalable_oversight_weak_llms_2024 |
On scalable oversight with weak LLMs judging strong LLMs | external_literature |
scalable_oversight |
scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control) | scalable-oversight-and-adversarial-ai-control | source | source note available; raw/cache text not published | Primary scalable-oversight comparator for protocol-specific weak-judge evaluations, debate and consultancy baselines, information-asymmetry limits, and open-role persuasion risks; it does not establish local judge calibration, debate efficacy, training safety, execution authority, or an ASI Stack result. |
viea |
Verified Intent-to-Execution Architecture | must_use |
whole_stack_execution_spine |
asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); system-boundaries-and-authority (System Boundaries and Authority); failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); human-intent-as-a-formal-input (Human Intent as a Formal Input); human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); stable-capability-fields (Stable Capability Fields); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); labor-os-and-typed-jobs (Labor OS and Typed Jobs); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); resource-economics-and-token-budgets (Resource Economics and Token Budgets); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology) | 01, 07, 12, 15, human-factors-and-meaningful-control-in-oversight, governed-operations-incident-command-and-graceful-degradation | source | source note available; exact source published in the live-book paper library | Keystone source. Human intent -> command contracts -> artifacts -> routing -> runtime targets -> verification -> deployment -> feedback. |
scf |
Stable Capability Fields | must_use |
governance_recursive_self_improvement |
asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); system-boundaries-and-authority (System Boundaries and Authority); failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); stable-capability-fields (Stable Capability Fields); capability-replacement-and-rollback (Capability Replacement and Rollback); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 05, replaceable-cognitive-substrates-beyond-transformer-monoculture, human-factors-and-meaningful-control-in-oversight, governed-operations-incident-command-and-graceful-degradation | source | source note available; exact source published in the live-book paper library | Use public release v1.0 when available. Stable boundaries, replacement, bounded authority, recoverable evolution. |
planforge |
PlanForge | must_use |
planning_control |
human-intent-as-a-formal-input (Human Intent as a Formal Input); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 06, governed-world-models-and-reality-grounding | source | source note available; exact source published in the live-book paper library | Planning substrate. Goal-to-execution compilation, hierarchical decomposition, DAG planning, scheduling, intelligence arbitrage. |
planforge_compiler_arch |
PlanForge: A Compiler Architecture for AI Task Orchestration | must_use_variant |
planning_control |
planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR) | 06 | source | source note available; exact source published in the live-book paper library | Later/alternate PlanForge framing. Prefer highest-quality/latest content after comparison. |
cognitive_compilation |
Cognitive Compilation | must_use |
planning_semantic_ir |
human-intent-as-a-formal-input (Human Intent as a Formal Input); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); mathematical-and-search-substrates (Mathematical and Search Substrates) | 06, 12, governed-world-models-and-reality-grounding | source | source note available; exact source published in the live-book paper library | Compiler framing for LLM-centered planning, semantic IR, target compilation, incremental repair. |
talos |
Talos Protocol | must_use |
labor_execution_os |
asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); system-boundaries-and-authority (System Boundaries and Authority); failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); human-intent-as-a-formal-input (Human Intent as a Formal Input); human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); human-ai-communication-persuasion-and-epistemic-security (Human-AI Communication, Persuasion, and Epistemic Security); societal-resilience-and-misuse-defense (Societal Resilience and Misuse Defense); stable-capability-fields (Stable Capability Fields); capability-replacement-and-rollback (Capability Replacement and Rollback); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); labor-os-and-typed-jobs (Labor OS and Typed Jobs); human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 08, inter-stack-protocols-identity-and-economic-exchange, human-factors-and-meaningful-control-in-oversight, governed-operations-incident-command-and-graceful-degradation, human-ai-communication-persuasion-and-epistemic-security | source | source note available; exact source published in the live-book paper library | AI labor OS. Deterministic cognitive manufacturing, typed jobs, control planes, auditability, tool isolation. |
talos_md |
Talos_Protocol_v1.0.md | must_use_variant |
labor_execution_os |
labor-os-and-typed-jobs (Labor OS and Typed Jobs) | 08 | source | source note available; connector-readable; raw text not published | Markdown/public release version. |
vcm_public |
Virtual_Context_Memory_v1 | must_use |
memory_context |
failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 07, inter-stack-protocols-identity-and-economic-exchange | source | source note available; exact source published in the live-book paper library | Public VCM release. Governed protocol for compiled working context. |
vcm_editable |
Virtual_Context_Memory_v1.0_Editable | must_use_variant |
memory_context |
failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 07 | source | source note available; connector-readable; raw text not published | Editable version with evidence-carrying planner-guided context compiler framing. |
spinoza |
Proof of Belief / The Spinoza Architecture | must_use |
reasoning_epistemology |
failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); fast-generation-architectures (Fast Generation Architectures); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 09 | source | source note available; exact source published in the live-book paper library | Neurosymbolic belief, transparent axiomatic AI belief systems, verification, belief revision. |
spinoza_composer |
Spinoza Composer / Spinoza Trinity | supporting |
reasoning_media_compliance |
artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay) | 09, 08 | source | source note available; exact source published in the live-book paper library | Narrative/compliance operating system variant. Use as applied Spinoza. |
moecot |
MoECOT-Agent Architecture Whitepaper | must_use |
implementation_reference |
asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); system-boundaries-and-authority (System Boundaries and Authority); failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); stable-capability-fields (Stable Capability Fields); capability-replacement-and-rollback (Capability Replacement and Rollback); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); labor-os-and-typed-jobs (Labor OS and Typed Jobs); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 10, 15, 16 | source | source note available; connector-readable; raw text not published | Full authenticated v1.1 architecture-whitepaper text passage-reviewed. Supports a compact explicit orchestrator, bounded specialist lanes, task state machines, leases/retries/dead-letter handling, fail-closed side-effect envelopes, run/task/control-plane ledgers, readiness distinct from routing, benchmark lanes, replay/handoff, provenance, architecture fingerprints, and governed improvement proposals. Source-reported runtime and benchmark artifacts were not imported or reproduced; the pinned MoECOT project dossier is the stronger implementation-reference record. |
moecot_md |
moecot_agent_whitepaper.md | must_use_variant |
implementation_reference |
routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology) | 10, 15, 16 | source | source note available; connector-readable; raw text not published | Full authenticated Markdown export reconciled with the primary MoECOT v1.1 Google Doc. It is a format and terminology variant, not independent corroboration or an additional empirical result. |
octopus_router |
Octopus Router Architecture | must_use |
routing_modular_intelligence |
routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); integrated-reference-architecture (Integrated Reference Architecture) | 10 | source | source note available; exact source published in the live-book paper library | Lightweight head/router with dynamically loaded specialist arms and local boundaries. |
rmi |
Ratcheting Modular Intelligence | must_use |
capability_ratchet |
the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); capability-replacement-and-rollback (Capability Replacement and Rollback); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); open-ended-improvement-engines (Open-Ended Improvement Engines); embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); fast-generation-architectures (Fast Generation Architectures); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 10, 13, replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; exact source published in the live-book paper library | Benchmark pressure, residual escrow, verified modular capability, regression preservation. |
cognitive_loop_closure |
Cognitive Loop Closure | must_use |
procedural_memory |
capability-replacement-and-rollback (Capability Replacement and Rollback); open-ended-improvement-engines (Open-Ended Improvement Engines); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); fast-generation-architectures (Fast Generation Architectures); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); living-book-methodology (Living Book Methodology) | 08, 10 | source | source note available; exact source published in the live-book paper library | Repeated cognition should become procedural memory / verified tools. |
benchmaxxing |
Benchmaxxing: The Performance Ratchet | must_use |
benchmarks_evidence |
dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); capability-replacement-and-rollback (Capability Replacement and Rollback); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); open-ended-improvement-engines (Open-Ended Improvement Engines); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); fast-generation-architectures (Fast Generation Architectures); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); capability-thresholds-and-deployment-commitments (Capability Thresholds and Deployment Commitments); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception); safety-cases-and-structured-assurance (Safety Cases and Structured Assurance); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 13, capability-thresholds-and-deployment-commitments, safety-cases-and-structured-assurance, replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; exact source published in the live-book paper library | Benchmarks as pressure surfaces, saturation -> regression, harder frontier, anti-Goodhart safeguards. |
cgs |
Compact Generative Systems | must_use |
compression_representation |
the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); fast-generation-architectures (Fast Generation Architectures); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression); resource-economics-and-token-budgets (Resource Economics and Token Budgets); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 11 | source | source note available; exact source published in the live-book paper library | Smallest adequate structure that can generate/govern target without hiding residual complexity. |
rgs |
Ratcheting Generative Systems | supporting |
compression_capability_growth |
procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) | 11, 10 | source | source note available; exact source published in the live-book paper library | Bridge between active compression, procedural memory, benchmark frontiers, verified AI growth. |
rankfold_neuralfold |
RankFold + NeuralFold | must_use |
compression_representation |
the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | 11 | source | source note available; exact source published in the live-book paper library | Tensor/artifact compression. Low-rank residual coding plus functional preprocessing and probe-route fallback. |
rankfold_compressor |
rankFold compressor | must_use_variant |
compression_representation |
rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression) | 11 | source | source note available; exact source published in the live-book paper library | Alternate RankFold/NeuralFold source. |
bbvca_v9 |
BBVCA_v9_final_public_release | must_use |
compression_representation |
the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression) | 11 | source | source note available; exact source published in the live-book paper library | Prefer v9. Generate-verify-repair compression from seeded local laws, bounded search, two-phase rate discipline. |
bbvca_main |
Big Bang Volumetric Compression Architecture | must_use_variant |
compression_representation |
compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty) | 11 | source | source note available; exact source published in the live-book paper library | Earlier/main BBVCA family doc. |
genesiscode |
GenesisCode | must_use |
executable_specification |
system-boundaries-and-authority (System Boundaries and Authority); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); labor-os-and-typed-jobs (Labor OS and Typed Jobs); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); mathematical-and-search-substrates (Mathematical and Search Substrates); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 08, 12 | source | source note available; exact source published in the live-book paper library | Tiny pure calculus + obligations + provenance for auditable AI-symbiotic programming. |
alignment_field |
Field of God / Alignment Field family | must_use |
alignment_constitution |
constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); inner-alignment-mesa-optimization-and-learned-objective-integrity (Inner Alignment, Mesa-Optimization, and Learned-Objective Integrity); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); governed-objective-formation-value-learning-and-goal-integrity (Governed Objective Formation, Value Learning, and Goal Integrity); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); resource-economics-and-token-budgets (Resource Economics and Token Budgets); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | constitutional-alignment-substrate, inner-alignment-mesa-optimization-and-learned-objective-integrity, moral-uncertainty-and-value-conflict, governed-objective-formation-value-learning-and-goal-integrity, security-kernel-and-digital-scifs, recursive-self-improvement-boundaries, resource-economics-and-token-budgets, integrated-reference-architecture, open-research-agenda-and-bibliography-plan | source | source note available; exact source published in the live-book paper library | Corben-authored long-form metaphysics, consciousness, ethics, AI-rights, and governance family. Its complete section-family audit preserves the five-factor consciousness heuristic only as theory-relative question decomposition; adds a prospective architecture-induced moral-risk review for self-preservation, persistent identity, valence-like state, and copy proliferation; separates copy, causal, memory, legal, authority, consent, and first-person continuity; and rejects the source’s scalar moral ranking, collective-consciousness, inevitable ethical convergence, physics, clinical, Omega, and karma claims as technical evidence. |
field_of_god |
The Field of God | must_use_variant |
alignment_constitution |
failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility) | constitutional-alignment-substrate, moral-uncertainty-and-value-conflict, governed-objective-formation-value-learning-and-goal-integrity, inner-alignment-mesa-optimization-and-learned-objective-integrity, human-ai-organizations-delegation-and-accountability, recursive-self-improvement-boundaries, integrated-reference-architecture | source | source note available; exact source published in the live-book paper library | Corben-authored predecessor to the Alignment Field family, containing title exploration, outline, abbreviated and expanded eight-part drafts across informational-relational metaphysics, a five-factor consciousness heuristic, attractor ethics, AI, copy continuity, genealogy, and conclusion. The complete audit treats alignment_field as the controlling successor, preserves power-care divergence, nested optimization, dissent/feedback, and identity-continuity distinctions, and explicitly rejects double counting, scalar moral ranking, collective consciousness, metaphysical proof, clinical/physics claims, upload survival, and teleology as evidence. |
field_of_god_ai_constitution |
Field of God AI Constitution | core_alignment_source |
constitutional_alignment_runtime_governance |
constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | constitutional-alignment-substrate, moral-uncertainty-and-value-conflict, recursive-self-improvement-boundaries, runtime-adapters | source | source note available; raw/cache text not published | Recovered in the Project Theseus repository. Constitutional alignment core for truth alignment, agency preservation, consent, non-domination, consciousness caution, least sufficient power, auditability, self-authorization limits, and runtime checks; use as source material only after source-note creation, not as proof or test evidence. |
ethica_mechanica |
Ethica Mechanica | must_use_variant |
alignment_constitution |
constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance) | constitutional-alignment-substrate, moral-uncertainty-and-value-conflict, governed-objective-formation-value-learning-and-goal-integrity, ai-deployment-transition-distribution-and-human-agency, privacy-data-rights-and-information-flow-governance, inter-stack-protocols-identity-and-economic-exchange, integrated-reference-architecture | source | source note available; exact source published in the live-book paper library | Corben-authored January 2026 philosophical and socio-technical treatise. Its complete audit retains the separation between bounded machine logistics and human normative authority, recursive feedback, dissent, governing-logic transparency versus personal privacy, distributional simulation as contestable evidence, and the requirement that exit/fork rights be materially exercisable under portability, network, compute, continuity, safety, privacy, and obligation constraints. Metaphysics, the consciousness equation, moral proofs, automatic veil legitimacy, unrestricted fork, and executable-protocol claims remain unsupported. |
eternal_code |
The Eternal Code / unified God, reality, conscious, alignment | must_use_variant |
alignment_constitution |
constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility) | constitutional-alignment-substrate, evidence-states-and-claim-discipline, spinoza-verification-and-proof-carrying-claims, tribunal-adversarial-review-and-claim-conflict, governed-world-models-and-reality-grounding, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval, moral-uncertainty-and-value-conflict, resource-economics-and-token-budgets | source | source note available; exact source published in the live-book paper library | Corben-authored multi-version computational-metaphysics and alignment family. Its complete audit traces the progression from an aggregate alignment score to Truth/Social/Task axes, a geometric product, heterarchical evaluator ownership, compute-aware exit, and a standing adversarial challenger. The book retains separate non-compensating epistemic, task, affected-party/constitutional, and authority/effect planes, evaluator lineage, and material exit costs while rejecting the consciousness and alignment formulas, oracle labels, consensus-as-truth, automatic energy throttling, fixed compute entitlement, theological/metaphysical claims, and moral proofs. |
coherence_exchange |
The Coherence Exchange | strong_support |
epistemic_market_synthesis |
evidence-states-and-claim-discipline (Evidence States and Claim Discipline); constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); ai-deployment-transition-distribution-and-human-agency (AI Deployment, Transition, Distribution, and Human Agency); multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); prototype-roadmap (Prototype Roadmap) | 09, 13, 15, institutions-international-coordination-and-public-legitimacy, ai-deployment-transition-distribution-and-human-agency | source | source note available; connector-readable; raw text not published | Found in AI generated paper dump. Use carefully; speculative synthesis of PlanForge, Spinoza, Talos, UAT, Alignment Field. |
verification_bandwidth |
Verification Bandwidth in Bounded Contexts | strong_support |
context_verification_theory |
evidence-states-and-claim-discipline (Evidence States and Claim Discipline); scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); fast-generation-architectures (Fast Generation Architectures); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | verification-bandwidth-and-context-adequacy, virtual-context-abi, evidence-states-and-claim-discipline, spinoza-verification-and-proof-carrying-claims, compact-generative-systems-and-residual-honesty, fast-generation-architectures, policy-optimization-and-learning-from-feedback, governed-deliberation-and-test-time-scaling, scalable-oversight-and-adversarial-ai-control, open-research-agenda-and-bibliography-plan | source | source note available; exact source published in the live-book paper library | Corben-authored version 1.0 context-verification hypothesis. The complete audit preserves the generation-versus-verification distinction, claim-relative semantic units and effective workspace, dominant-component pressure, explicit interaction obligations, decomposition boundaries, a coherency-horizon escalation rule, and a stronger held-out contradiction protocol. It rejects the four named ‘theorems’ as proved laws: dense joint attention is neither necessary nor sufficient, lossy compression need not discard property-relevant information, DPI does not establish monotonic LLM contradiction growth, all-pair checking is not universally required, uniform half-window partitioning is not generally optimal, and RAG may retrieve exact text. |
beastbrain |
BeastBrain Cognitive Architecture | supporting_lineage |
whole_stack_lineage |
asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); prototype-roadmap (Prototype Roadmap) | asi-is-a-stack-not-a-model, the-efficient-asi-hypothesis, routing-heads-and-specialist-cores, personal-compute-hives-and-federated-edge-intelligence, governed-model-training-distributed-optimization-and-scaling, fast-generation-architectures, security-kernel-and-digital-scifs, perception-sensor-fusion-and-observation-trust, prototype-roadmap, integrated-reference-architecture | source | source note available; exact source published in the live-book paper library | Corben-authored 70,000-word evolving architecture notebook spanning early blueprints through versions 1.0–6.1. The complete family audit preserves whole-system/homeostatic design, versioned hardware-profile qualification, physical memory-tier and residency accounting, distinct memory forms, governed consolidation and retention, substrate-neutral routing, contract-first planning and implementation, dependency-aware parallelism, opaque secret handles, multimodal perception, distributed service and artifact interfaces, and maintenance-learning windows. It rejects master-label maturity, repeated-version double counting, infinite-context/zero-copy/power/performance projections, geometric-truth and ignorance theorems, scalar routing/retention authority, unsafe forced self-evolution, test-suite sufficiency, tribunal consensus, software-SCIF guarantees, SSD-erasure assumptions, censorship-resistance, and autonomous update claims. |
beastbrain_timeless |
BeastBrain Architecture: Timeless Edition | supporting_lineage |
whole_stack_lineage |
the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); prototype-roadmap (Prototype Roadmap) | the-efficient-asi-hypothesis, personal-compute-hives-and-federated-edge-intelligence, prototype-roadmap | source | source note available; exact source published in the live-book paper library | Corben-authored standalone export of BeastBrain v3.3.4. The complete audit treats it as a near-duplicate Timeless branch already embedded in the main BeastBrain corpus, not independent support. It preserves evergreen whole-stack and Mimic hardware-adaptation framing while explicitly rejecting forced self-evolution, geometric truth, fixed entropy routing, infinite context, zero-copy, power, leak-resistance, constant-time verification, hardware-adaptation, and distributed-scaling projections as evidence. |
aletheia |
Aletheia / Proof-Carrying Workbench Lineage | supporting_lineage |
safe_general_intelligence_lineage |
asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); scientific-discovery-and-experimental-governance (Scientific Discovery and Experimental Governance) | asi-is-a-stack-not-a-model, the-efficient-asi-hypothesis, claim-ledgers-and-belief-revision, scientific-discovery-and-experimental-governance | source | source note available; exact source published in the live-book paper library | Complete four-version correction-lineage audit from the original Aletheia epistemic-engine proposal through Aletheia Foundry and Proof-Carrying Workbench v1.1/v1.2. The later PCW design controls conflicts: semantic scope rather than contract-hash theater, claim-native release surfaces, separate assurance classes, least-privilege capabilities, bounded adversarial review, governed commitments, template decay, recertification, and incident response. The audit rejects immutable primitives, scalar truth/intervention scores, live-oracle and consensus truth, deterministic open-domain claim extraction, universal source allowlists, unvalidated risk thresholds, and safe-general-intelligence claims. |
context_engineer |
Context Engineer / Manhattan Protocol | supporting_lineage |
memory_context_lineage |
security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint) | virtual-context-abi, context-transactions-snapshots-mounts-and-taint, security-kernel-and-digital-scifs | source | source note available; exact source published in the live-book paper library | Complete two-version audit of the Manhattan Protocol context-supply-chain paper. Preserves the context governor, layered representations, mission briefs, proposed MCP memory fields, need-to-know admission, and compartment lifecycle while separating sanitization, declassification, memory commit, zeroization, revocation, and residuals. Rejects Ring Attention as physical isolation, protocol fields as enforcement, regex/entropy scanning as semantic non-disclosure, permanent-wipe language, infinite storage, and all unreproduced benchmark figures. |
black_hole_context_manager |
Black Hole Context Manager | supporting_lineage |
memory_context_lineage |
context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint) | context-transactions-snapshots-mounts-and-taint | source | source note available; exact source published in the live-book paper library | Complete v5.0/v4.1 pseudocode-lineage audit. Preserves versioned context units, tiered placement, lazy task-relative evaluation, explicit goal-drift decisions, reversible freeze/thaw with hysteresis, protected low-entropy constraints, and factual-retrieval versus generative-reconstruction separation. Treats character entropy, semantic mass, K-means thresholds, HMAC, repeated confirmation, keyword routing, and the Drifting Needle as fallible candidates or weak baselines; rejects production-ready and security claims. |
ladon_manhattan |
Ladon & The Manhattan Protocol | supporting_lineage |
security_governance |
system-boundaries-and-authority (System Boundaries and Authority); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); stable-capability-fields (Stable Capability Fields); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) | security-kernel-and-digital-scifs, runtime-adapters-tool-permissions-and-human-approval, system-boundaries-and-authority | source | source note available; exact source published in the live-book paper library | Complete standalone security-paper audit. Preserves capability/credential separation, opaque caller-bound handles, trusted input paths, late substitution or remote use, per-use policy checks, and an explicit compartment lifecycle. The audit separates secret non-disclosure from authority misuse and harmful effects, treats returned artifacts as possible sensitive derivatives, and rejects platform equivalence, PROT_NONE/enclave conflation, portable trusted-UI claims, incomplete Rust pseudocode as implementation, and the paper’s Ignorance and Ephemerality ‘theorems’. |
uat |
Unified Adaptive Tribunal | supporting_lineage |
evaluation_refinement |
evidence-states-and-claim-discipline (Evidence States and Claim Discipline); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) | spinoza-verification-and-proof-carrying-claims, claim-ledgers-and-belief-revision, evidence-states-and-claim-discipline, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; exact source published in the live-book paper library | Complete three-tab correction-lineage audit. The public human-in-the-loop architecture controls the original promotional multi-model tribunal: it preserves structural/retrieval/dialectical candidate views, an explicit dossier boundary and omitted frontier, probabilistic claim extraction, richer proposition states, bounded adversarial review, compression fidelity, and accountable human handoffs. It rejects brand-count diversity, consensus and stability as truth, delete-by-dossier-absence, SVO completeness, fixed thresholds, guard-model authority, all claimed performance/cost figures, and production or superiority labels. |
treellm |
TreeLLM | supporting_lineage |
semantic_representation |
cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); mathematical-and-search-substrates (Mathematical and Search Substrates) | 09, 11 | source | source note available; exact source published in the live-book paper library | Hierarchical semantic token system for grounded, efficient, explainable language modeling. |
software_magic_grimoire |
The Grimoire of Software Magic Words: Operative Vocabulary, Prompt-Spells, and Stacked Workflows | supporting_lineage |
command_contracts_promptcraft |
human-intent-as-a-formal-input (Human Intent as a Formal Input); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); labor-os-and-typed-jobs (Labor OS and Typed Jobs); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) | 06, 08 | source | source note available; exact source published in the live-book paper library | Full composite-document audit completed 2026-07-31 across the public grimoire, 1,645-entry lexicon, pocket edition, stacked-spells addendum, and prompt pack. Retains bounded instruction fields, layered identity, typed handoffs, guards, evidence loops, scoped recursion, recovery, and workflow versioning; rejects vocabulary, role prompts, Gödel numbers, coil geometry, templates, and authored examples as evidence of meaning, authority, performance, or safety. |
road_to_agi |
Road To AGI | supporting_lineage |
strategic_roadmap |
benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 16 | source | source note available; connector-readable; raw text not published | Remaining work / roadmap context. |
simulation_scaling |
The Simulation Scaling Law: Resource Constraints on Scope, Clockspeed, and Effective Fidelity in Nested Physical Simulations | optional_support |
compute_fidelity_constraints |
the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); resource-economics-and-token-budgets (Resource Economics and Token Budgets); mathematical-and-search-substrates (Mathematical and Search Substrates); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 03, 11, appendix, governed-world-models-and-reality-grounding, learning-theory-generalization-and-scaling-science | source | source note available; exact source published in the live-book paper library | Full six-tab lineage audit completed 2026-07-31. Retains the prospective simulation contract, typed bottleneck accounting, and logical-possibility/physical-feasibility split; treats D = scope*clockspeed/efficiency <= capacity as a conditional scalar heuristic rather than a proved universal law, rejects unsupported 1:1 and physical-limit overclaims, and routes simulator adequacy and transfer through Resource Economics. |
tokenmana |
TokenMana | optional_support |
resource_economics |
planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets); physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) | resource-economics-and-token-budgets, human-factors-and-meaningful-control-in-oversight, benchmark-ratchets-and-anti-goodhart-evidence, physical-compute-infrastructure-energy-and-environmental-constraints | source | source note available; exact source published in the live-book paper library | Complete five-tab/two-paper correction-lineage audit. Preserves regenerative capacity as one candidate budget mechanism and adds temporal-access contracts that expose renewal, accrual, expiry, burst, pricing, notification, fairness, and human-schedule effects. Separates renewal clustering, nocturnal work, sleep, and cognitive-friction proxies; strengthens privacy against employment/medical inference; and records mathematical gaps in stock units/boundaries, equilibrium, control continuity, strict variance, feedback stability, and profit claims. No theorem, simulation, load result, pricing result, human outcome, or welfare result is promoted. |
coilmoecot |
CoilMoECOT Whitepaper v2.0 | optional_technical_appendix |
mathematical_search_substrate |
routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); mathematical-and-search-substrates (Mathematical and Search Substrates); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | 12 | source | source note available; connector-readable; raw text not published | Full composite connector text section-audited. Retains prime-temporal trace features, Graph/Trace-first placement, ledger-derived diagnostic tuples, removable shadow lanes, explicit pre-planner/post-plan/post-run insertion points, anti-experts as visible penalty signals, bounded update slices, and benchmark/canary/rollback promotion. Repeated MoECOT, packaging, and task-system appendices are treated as shared substrate rather than independent evidence. |
temporal_coil_research |
Temporal Coil Research | optional_technical_appendix |
mathematical_search_substrate |
mathematical-and-search-substrates (Mathematical and Search Substrates); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) | 12 | source | source note available; raw/cache text not published | Full experiment note reviewed. Reports 11 variants, three seeds, six rounds per variant, split winner frequency, small mean deltas, flat pass/reward/holdout lanes, separation dominated by the collapse composite, and one narrow threshold-tuned seed. Preserved as an inconclusive source-reported result and placement-confounding lesson, not proof of benefit or general failure. |
bugbrain |
BugBrain / Project Genesis: Neuro-Symbolic Bare-Metal Edge Intelligence Paper Lineage | optional_support |
edge_efficiency_lineage |
compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology) | 11, 16, appendix | source | source note available; exact source published in the live-book paper library | Full 15-tab paper-lineage audit completed 2026-07-31 and reconciled against the pinned implementation dossier. Retains hardware-explicit ownership, state-qualified capacity, compact typed graphs, tiered context/paging/persistence, one-shot authority, artifact replay, readiness semantics, and objective-term effect tests; rejects consciousness, AGI, completeness, projected performance, named-module, source-presence, and skipped-green claims that outrun implementation evidence. |
cca_project |
Compiled Cognitive Architecture project | implementation_reference |
compiled_cognitive_architecture |
system-boundaries-and-authority (System Boundaries and Authority); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) | evidence-states-and-claim-discipline, system-boundaries-and-authority, artifact-graphs-audit-logs-and-replay, cognitive-compilation-and-semantic-ir, virtual-context-abi, context-transactions-snapshots-mounts-and-taint, claim-ledgers-and-belief-revision, spinoza-verification-and-proof-carrying-claims, procedural-memory-and-cognitive-loop-closure, ai-supply-chain-integrity-and-lifecycle-provenance, model-weight-custody-and-hardware-roots-of-trust, recursive-self-improvement-boundaries, governed-deliberation-and-test-time-scaling, benchmark-ratchets-and-anti-goodhart-evidence, readiness-gates-residual-escrow-and-quarantine, executable-specifications-and-lean-proof-envelope, personal-compute-hives-and-federated-edge-intelligence, intent-to-execution-contracts, integrated-reference-architecture | local project reference (not publicly linked) | source note available; raw/cache text not published | Pinned local-project convergence reference for external semantic memory, semantic compilation, epistemic governance, bounded self-modification, proof/runtime coupling, benchmark truth, and negative transfer evidence; public-safe source note only, with no reproduced capability or safety result. |
moecot_manifest_project |
MoECOT Manifest compiler-era project | implementation_reference |
compiler_first_ai_systems |
system-boundaries-and-authority (System Boundaries and Authority); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) | evidence-states-and-claim-discipline, system-boundaries-and-authority, security-kernel-and-digital-scifs, integrated-reference-architecture, cognitive-compilation-and-semantic-ir, artifact-graphs-audit-logs-and-replay, ai-supply-chain-integrity-and-lifecycle-provenance, routing-heads-and-specialist-cores, open-ended-improvement-engines, recursive-self-improvement-boundaries, benchmark-ratchets-and-anti-goodhart-evidence, readiness-gates-residual-escrow-and-quarantine, executable-specifications-and-lean-proof-envelope, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval, virtual-context-abi, claim-ledgers-and-belief-revision, spinoza-verification-and-proof-carrying-claims, procedural-memory-and-cognitive-loop-closure, artifact-steward-agents-and-living-project-governance | local project reference (not publicly linked) | source note available; raw/cache text not published | Pinned local-project compiler/control-plane reference for semantic IR, manifest compilation, registry truth, context and memory governance, target portability, provenance, multi-agent training, bounded self-improvement, and the negative gap between internal contracts and external holdout capability; public-safe source note only, with no reproduced capability or safety result. |
beastbrain_project |
BeastBrain historical AI system project | implementation_reference |
durable_semantic_memory_and_system_architecture |
system-boundaries-and-authority (System Boundaries and Authority); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) | evidence-states-and-claim-discipline, system-boundaries-and-authority, routing-heads-and-specialist-cores, planning-as-a-control-layer, virtual-context-abi, claim-ledgers-and-belief-revision, spinoza-verification-and-proof-carrying-claims, procedural-memory-and-cognitive-loop-closure, resource-economics-and-token-budgets, benchmark-ratchets-and-anti-goodhart-evidence, readiness-gates-residual-escrow-and-quarantine, security-kernel-and-digital-scifs, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval, personal-compute-hives-and-federated-edge-intelligence, artifact-steward-agents-and-living-project-governance | local project reference (not publicly linked) | source note available; raw/cache text not published | Hashed local-project primary lineage for DKL, Portia, semantic coordinates, bounded knowledge snapshots, PlanForge, SSD-first memory, and organism-style cognition, plus negative implementation evidence on simulations, router effects, security handles, tribunal stubs, ontology drift, compile history, and readiness overclaim; public-safe source note only, with no reproduced capability or safety result. |
bugbrain_project |
BugBrain bare-metal neuro-symbolic intelligence project | implementation_reference |
hardware_explicit_governed_cognition |
system-boundaries-and-authority (System Boundaries and Authority); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) | evidence-states-and-claim-discipline, system-boundaries-and-authority, model-weight-custody-and-hardware-roots-of-trust, security-kernel-and-digital-scifs, virtual-context-abi, context-transactions-snapshots-mounts-and-taint, claim-ledgers-and-belief-revision, spinoza-verification-and-proof-carrying-claims, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval, artifact-graphs-audit-logs-and-replay, integrated-reference-architecture, resource-economics-and-token-budgets, readiness-gates-residual-escrow-and-quarantine, benchmark-ratchets-and-anti-goodhart-evidence, ai-supply-chain-integrity-and-lifecycle-provenance | local project reference (not publicly linked) | source note available; raw/cache text not published | Pinned local-project bare-metal implementation reference for explicit core ownership, compact typed graphs, signed tiered context, privileged-action lifecycle, protocol security, artifact replay, and readiness, plus negative evidence on capacity/residency arithmetic, theory-labelled proxies, fixed random cognitive modules, root-of-trust assumptions, skipped checks, and narrative/report divergence; public-safe source note only, with no reproduced hardware capability or safety result. |
corbens_trainer_project |
Corben’s Trainer epistemic training and evaluation control plane | implementation_reference |
epistemic_training_and_evaluation_control_plane |
system-boundaries-and-authority (System Boundaries and Authority); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) | ai-supply-chain-integrity-and-lifecycle-provenance, artifact-graphs-audit-logs-and-replay, artifact-steward-agents-and-living-project-governance, benchmark-ratchets-and-anti-goodhart-evidence, evidence-states-and-claim-discipline, governed-model-training-distributed-optimization-and-scaling, integrated-reference-architecture, open-ended-improvement-engines, procedural-memory-and-cognitive-loop-closure, readiness-gates-residual-escrow-and-quarantine, recursive-self-improvement-boundaries, resource-economics-and-token-budgets, runtime-adapters-tool-permissions-and-human-approval, security-kernel-and-digital-scifs, system-boundaries-and-authority | local project reference (not publicly linked) | source note available; raw/cache text not published | Pinned local-project implementation reference for typed experiment manifests, trainer/backend separation, artifact lineage, learning-truth gates, benchmark authenticity, quarantine, claim derivation, and revocable promotion boundaries, plus negative implementation evidence on seed identity, content pinning, decontamination, transitive revocation, checkpoint acknowledgement, and report divergence; public-safe source note only, with no reproduced model capability or safety result. |
corbens_best_model_possible_project |
Corben’s Best Model Possible recurrent-model and mechanism laboratory | implementation_reference |
recurrent_model_mechanisms_and_capability_evidence |
system-boundaries-and-authority (System Boundaries and Authority); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) | routing-heads-and-specialist-cores, system-boundaries-and-authority, security-kernel-and-digital-scifs, governed-deliberation-and-test-time-scaling, cognitive-compilation-and-semantic-ir, integrated-reference-architecture, open-ended-improvement-engines, recursive-self-improvement-boundaries, evidence-states-and-claim-discipline, benchmark-ratchets-and-anti-goodhart-evidence, artifact-graphs-audit-logs-and-replay, ai-supply-chain-integrity-and-lifecycle-provenance, claim-ledgers-and-belief-revision, spinoza-verification-and-proof-carrying-claims, context-transactions-snapshots-mounts-and-taint, runtime-adapters-tool-permissions-and-human-approval, readiness-gates-residual-escrow-and-quarantine, executable-specifications-and-lean-proof-envelope, intent-to-execution-contracts | local project reference (not publicly linked) | source note available; raw/cache text not published | Pinned local-project implementation and negative-case reference for shared-weight recurrence, fixed-feature adapter learning, specialist checkpoint banks, architecture-search discipline, semantic compilation, memory, routing, governance, verification, tools, speech, and metric provenance; public-safe source note only, with no reproduced trained-foundation-model, general-generation, external-capability, or safety result. |
project_theseus_whitepaper |
Project Theseus Whitepaper | implementation_reference |
report_first_rmi_prototype |
procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | implementation, prototype, appendix | source | source note available; raw/cache text not published | Local-first report-driven RMI implementation reference: SymLiquid, SparkStream, Octopus Router, residual escrow, self-evolution gates, Hive runtime, observability. |
theseus_plan_compiler |
Theseus Plan Compiler | implementation_reference |
planning_control |
integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap) | planning, execution, prototype | source | source note available; raw/cache text not published | Goal-to-contract compiler with semantic IR DAGs, VCM context slices, executor routes, claim/evidence targets, contract hashes, and replay traces. |
theseus_self_evolution_system |
Theseus Self-Evolution System | implementation_reference |
recursive_self_improvement_governance |
recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap) | self-improvement, benchmarking, prototype | source | source note available; raw/cache text not published | Evidence-first self-evolution lane with intervention ladder, ATTD repo-health gate, guarded teacher self-edit, architecture experiment governance, loop closure, and outcome ledger. |
theseus_architecture_gate |
Theseus Architecture Gate | implementation_reference |
readiness_gate_governance |
recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); capability-thresholds-and-deployment-commitments (Capability Thresholds and Deployment Commitments); integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap) | governance, benchmarking, prototype, capability-thresholds-and-deployment-commitments | source | source note available; raw/cache text not published | Pre-training readiness gate covering ratchet completeness, router readiness, safety ledger, residual escrow, bridge benchmarks, procedural tools, routing memory, lifecycle governance, and external-inference zero. |
theseus_operator_os |
Hive Operator OS and Work Board | implementation_reference |
labor_os_operator_surface |
human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap) | execution, runtime, prototype, human-factors-and-meaningful-control-in-oversight, governed-operations-incident-command-and-graceful-degradation | source | source note available; raw/cache text not published | Shared command vocabulary, durable SQLite work board, node registry, background/watch/wake contracts, skill registry, tool hooks, feedback routing, and safety-visible operator surface. |
theseus_circle_transfer |
Theseus Circle Calculus Transfer Lane | implementation_reference |
proof_contract_transfer |
mathematical-and-search-substrates (Mathematical and Search Substrates); coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts); coilra-multicoil-rope-and-cyclic-mixers (CoilRA, MultiCoil RoPE, and Cyclic Mixers); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | math-substrates, proof-contracts, prototype | source | source note available; raw/cache text not published | Report-only bridge from Circle finite fixtures into private Theseus benchmark design with explicit quality/runtime/memory/transfer/failure-case claim boundaries. |
circle_calculus_core |
Circle Calculus | core_technical_source |
proof_carrying_mathematical_substrate |
mathematical-and-search-substrates (Mathematical and Search Substrates); circle-calculus-and-proof-carrying-ai-contracts (Circle Calculus and Proof-Carrying AI Contracts); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | math-substrates, proofs, appendix | source | source note available; raw/cache text not published | Proof-carrying finite cyclic mathematics project with Lean proofs, Python reference models, Rust utilities, theorem manifests, papers, and Quarto living book. |
circle_ai_contract_suite |
Circle Calculus AI Contract Suite | core_technical_source |
proof_carrying_ai_contracts |
circle-calculus-and-proof-carrying-ai-contracts (Circle Calculus and Proof-Carrying AI Contracts); coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts); coilra-multicoil-rope-and-cyclic-mixers (CoilRA, MultiCoil RoPE, and Cyclic Mixers); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | proof-contracts, attention-memory, appendix | source | source note available; raw/cache text not published | Theorem-linked AI contract families for RoPE, KV-cache freshness, sparse attention, recurrence schedules, strided fanout, cyclic memory, multicoil phase, cyclic mixers, and seed-rule regeneration. |
circle_ai_architectures |
Circle AI Architectures | core_technical_source |
cyclic_ai_architecture |
compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); mathematical-and-search-substrates (Mathematical and Search Substrates); circle-calculus-and-proof-carrying-ai-contracts (Circle Calculus and Proof-Carrying AI Contracts); coilra-multicoil-rope-and-cyclic-mixers (CoilRA, MultiCoil RoPE, and Cyclic Mixers) | math-substrates, semantic-representation | source | source note available; raw/cache text not published | Disciplined Circle AI thesis: use phase, recurrence, rotation, sparse cyclic mixing, circular memory, harmonic transforms, or geometry-aware structure only where the structure is real and baselines support it. |
coil_attention_memory |
Coil Attention and Memory | core_technical_source |
cyclic_attention_memory |
coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts) | memory, attention, recurrence | source | source note available; raw/cache text not published | Proof-linked cyclic memory, KV-cache freshness, sparse-attention coverage, recurrence schedules, loop-exit certificates, work budgets, and alias diagnostics. |
coilra_multicoil_rope |
CoilRA and MultiCoil RoPE | core_technical_source |
cyclic_mixers_position_encoding |
compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); resource-economics-and-token-budgets (Resource Economics and Token Budgets); coilra-multicoil-rope-and-cyclic-mixers (CoilRA, MultiCoil RoPE, and Cyclic Mixers) | representation, routing, resource-economics | source | source note available; raw/cache text not published | Adapter-block, residue/winding, block-cyclic, multicoil, relative RoPE, circulant convolution, cyclic mixer, and parameter-accounting substrate with explicit non-claims. |
rope_position_certifier |
Proof-Carrying RoPE Position Distinguishability | core_technical_source |
proof_carrying_position_contract |
circle-calculus-and-proof-carrying-ai-contracts (Circle Calculus and Proof-Carrying AI Contracts); coilra-multicoil-rope-and-cyclic-mixers (CoilRA, MultiCoil RoPE, and Cyclic Mixers); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope) | proof-contracts, representation | source | source note available; raw/cache text not published | Externally usable RoPE position-distinguishability certifier with theorem-linked exact collision reports, bounded real-phase frontier, machine-readable receipts, and explicit non-claims. |
proof_carrying_circular_computation |
Proof-Carrying Circular Computation | supporting_technical_source |
proof_carrying_compute_substrate |
mathematical-and-search-substrates (Mathematical and Search Substrates); circle-calculus-and-proof-carrying-ai-contracts (Circle Calculus and Proof-Carrying AI Contracts); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope) | proofs, runtime, math-substrates | source | source note available; raw/cache text not published | CoilIR-style path from circle/coil expressions to dictionary-recognized cyclic structure, Lean-proved rewrite/address transformations, backend selection, and benchmark validation. |
ext_concrete_ai_safety_2016 |
Concrete Problems in AI Safety | external_literature |
alignment_control |
failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence) | failure-modes-of-ungoverned-intelligence, evidence-states-and-claim-discipline, benchmark-ratchets-and-anti-goodhart-evidence, policy-optimization-and-learning-from-feedback | source | source note available; raw/cache text not published | External alignment/control source for accident-risk taxonomy: side effects, reward hacking, scalable supervision, safe exploration, and distributional shift. |
ext_goal_misgeneralization_2022 |
Goal Misgeneralization in Deep Reinforcement Learning | external_literature |
alignment_control |
failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); inner-alignment-mesa-optimization-and-learned-objective-integrity (Inner Alignment, Mesa-Optimization, and Learned-Objective Integrity); governed-objective-formation-value-learning-and-goal-integrity (Governed Objective Formation, Value Learning, and Goal Integrity) | failure-modes-of-ungoverned-intelligence, policy-optimization-and-learning-from-feedback, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | External alignment-control source for distinguishing capability generalization from goal generalization failures, used to ground goal-misbinding and out-of-distribution objective failure language. |
ext_learned_optimization_risks_2019 |
Risks from Learned Optimization in Advanced Machine Learning Systems | external_literature |
alignment_control |
failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); inner-alignment-mesa-optimization-and-learned-objective-integrity (Inner Alignment, Mesa-Optimization, and Learned-Objective Integrity); governed-objective-formation-value-learning-and-goal-integrity (Governed Objective Formation, Value Learning, and Goal Integrity) | failure-modes-of-ungoverned-intelligence, recursive-self-improvement-boundaries, policy-optimization-and-learning-from-feedback | source | source note available; raw/cache text not published | External alignment-control source for mesa-optimization and learned-objective mismatch, used to ground hidden optimizer, proxy-objective, and deceptive-alignment-adjacent failure language. |
ext_constitutional_ai_2022 |
Constitutional AI: Harmlessness from AI Feedback | external_literature |
alignment_control |
constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility) | constitutional-alignment-substrate, policy-optimization-and-learning-from-feedback | source | source note available; raw/cache text not published | External constitutional-AI source for training harmless assistants from a rule/principle list through supervised revision and AI-feedback reinforcement learning, used as a comparator for operational constitutional predicates. |
ext_collective_constitutional_ai_2024 |
Collective Constitutional AI: Aligning a Language Model with Public Input | external_literature |
alignment_governance |
constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance) | constitutional-alignment-substrate, moral-uncertainty-and-value-conflict | source | source note available; raw/cache text not published | External constitutional-AI governance source for sourcing and integrating public input into language-model principles, used as a comparator for constitution authorship, public input, contestability, and governance boundaries. |
ext_corrigibility_2015 |
Corrigibility | external_literature |
alignment_control |
constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); capability-replacement-and-rollback (Capability Replacement and Rollback) | constitutional-alignment-substrate, moral-uncertainty-and-value-conflict, recursive-self-improvement-boundaries, capability-replacement-and-rollback | source | source note available; raw/cache text not published | External corrigibility source for intervention tolerance, shutdown behavior, anti-manipulation incentives, and propagation across subsystems or self-modification. |
ext_off_switch_game_2016 |
The Off-Switch Game | external_literature |
alignment_control |
constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance) | constitutional-alignment-substrate, recursive-self-improvement-boundaries, runtime-adapters-tool-permissions-and-human-approval, moral-uncertainty-and-value-conflict | source | source note available; raw/cache text not published | External alignment source for shutdown incentives, uncertainty about objectives, and preserving human correction authority. |
ext_reinforcement_learning_moral_uncertainty_2020 |
Reinforcement Learning Under Moral Uncertainty | external_literature |
alignment_control |
moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance) | moral-uncertainty-and-value-conflict, policy-optimization-and-learning-from-feedback | source | source note available; raw/cache text not published | External AI moral-uncertainty source for agents acting under disagreement across moral theories, used as a comparator for value-conflict records and reward-function caveats. |
ext_contestable_ai_design_2022 |
Contestable AI by Design: Towards a Framework | external_literature |
governance_evals |
moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review) | moral-uncertainty-and-value-conflict, spinoza-verification-and-proof-carrying-claims | source | source note available; raw/cache text not published | External contestable-AI source for designing systems whose outcomes can be challenged, used as a comparator for dissent, appeal, audit, contestability, and governance-interface design. |
ext_optimal_policies_power_2019 |
Optimal Policies Tend to Seek Power | external_literature |
alignment_control |
failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); inner-alignment-mesa-optimization-and-learned-objective-integrity (Inner Alignment, Mesa-Optimization, and Learned-Objective Integrity) | failure-modes-of-ungoverned-intelligence, system-boundaries-and-authority, recursive-self-improvement-boundaries, readiness-gates-residual-escrow-and-quarantine | source | source note available; raw/cache text not published | External power-seeking source for formal analysis of option preservation and power-seeking tendencies under classes of reward functions and environments. |
ext_model_evaluation_extreme_risks_2023 |
Model evaluation for extreme risks | external_literature |
governance_evals |
dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); prototype-roadmap (Prototype Roadmap) | dangerous-capability-domains-and-misuse-uplift, benchmark-ratchets-and-anti-goodhart-evidence, readiness-gates-residual-escrow-and-quarantine, recursive-self-improvement-boundaries, prototype-roadmap | source | source note available; raw/cache text not published | External governance/evals source for dangerous capability evaluations, alignment evaluations, and deployment/security decisions under extreme-risk framing. |
ext_frontier_ai_regulation_2023 |
Frontier AI Regulation: Managing Emerging Risks to Public Safety | external_literature |
governance_evals |
living-book-methodology (Living Book Methodology) | system-boundaries-and-authority, readiness-gates-residual-escrow-and-quarantine, benchmark-ratchets-and-anti-goodhart-evidence, living-book-methodology | source | source note available; raw/cache text not published | External governance source for frontier AI standard setting, registration/reporting, compliance mechanisms, pre-deployment risk assessment, external scrutiny, and post-deployment monitoring. |
ext_nist_ai_rmf_1_0_2023 |
Artificial Intelligence Risk Management Framework (AI RMF 1.0) | external_literature |
governance_evals |
human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology) | system-boundaries-and-authority, readiness-gates-residual-escrow-and-quarantine, living-book-methodology, prototype-roadmap, governed-operations-incident-command-and-graceful-degradation | source | source note available; raw/cache text not published | Official NIST AI RMF 1.0 source for risk framing, trustworthiness characteristics, lifecycle roles, and Govern/Map/Measure/Manage functions. |
ext_owasp_llm_top_10_2025 |
OWASP Top 10 for LLMs and Gen AI Apps | external_literature |
ai_security |
security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs) | security-kernel-and-digital-scifs, runtime-adapters-tool-permissions-and-human-approval, artifact-steward-agents-and-living-project-governance | source | source note available; raw/cache text not published | Official OWASP GenAI security reference for LLM prompt injection, sensitive information disclosure, excessive agency, and related application-security risks. |
ext_nist_zero_trust_architecture_2020 |
Zero Trust Architecture | external_literature |
security_governance |
security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs) | security-kernel-and-digital-scifs, system-boundaries-and-authority, runtime-adapters-tool-permissions-and-human-approval | source | source note available; raw/cache text not published | Official NIST zero-trust architecture source for resource-centric access mediation, least-privilege access, policy enforcement points, and continuous authorization framing. |
ext_saltzer_schroeder_protection_1975 |
The Protection of Information in Computer Systems | external_literature |
security_principles |
security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs) | security-kernel-and-digital-scifs, system-boundaries-and-authority | source | source note available; raw/cache text not published | Classic security-principles source for least privilege, complete mediation, economy of mechanism, fail-safe defaults, separation of privilege, and open design as comparators for kernel-like AI security boundaries. |
ext_capability_based_computer_systems_1984 |
Capability-Based Computer Systems | external_literature |
capability_security |
stable-capability-fields (Stable Capability Fields) | system-boundaries-and-authority, stable-capability-fields | source | source note available; raw/cache text not published | External capability-system comparator for authority-bearing capabilities, protection domains, and permission boundaries that help position System Boundaries authority records and SCF authority ceilings without claiming ASI Stack capability enforcement. |
ext_confused_deputy_hardy_1988 |
The Confused Deputy: (or why capabilities might have been invented) | external_literature |
capability_security |
Unassigned in current structure | system-boundaries-and-authority, runtime-adapters-tool-permissions-and-human-approval | source | source note available; raw/cache text not published | External confused-deputy source for authority laundering, ambient authority, and the capability-security motivation for binding designation to permission at tool and handoff boundaries. |
ext_semver_2_0_0 |
Semantic Versioning 2.0.0 | external_literature |
interface_versioning |
stable-capability-fields (Stable Capability Fields) | stable-capability-fields | source | source note available; raw/cache text not published | External versioned-interface comparator for public API contracts, compatibility, and breaking-change signaling as a narrow baseline for SCF field versions and stable interfaces. |
ext_slsa_v1_0 |
SLSA v1.0 | external_literature |
supply_chain_provenance |
stable-capability-fields (Stable Capability Fields) | stable-capability-fields | source | source note available; raw/cache text not published | External supply-chain provenance comparator for artifact integrity, provenance, build levels, and dependency on verifiable artifacts before promotion or default route use. |
ext_react_2022 |
ReAct: Synergizing Reasoning and Acting in Language Models | external_literature |
planning_agent_control |
Unassigned in current structure | planning-as-a-control-layer, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval, artifact-graphs-audit-logs-and-replay | source | source note available; raw/cache text not published | External planning/agent-control source for interleaving reasoning traces with task-specific actions and environment or knowledge-base interaction. |
ext_tree_of_thoughts_2023 |
Tree of Thoughts: Deliberate Problem Solving with Large Language Models | external_literature |
planning_search |
cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling) | planning-as-a-control-layer, cognitive-compilation-and-semantic-ir, benchmark-ratchets-and-anti-goodhart-evidence, governed-deliberation-and-test-time-scaling | source | source note available; raw/cache text not published | External planning/search source for exploring, evaluating, and backtracking over multiple reasoning paths rather than left-to-right token continuation alone. |
ext_pddl_1998 |
PDDL: The Planning Domain Definition Language | external_literature |
planning_modeling |
cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR) | planning-as-a-control-layer, intent-to-execution-contracts, cognitive-compilation-and-semantic-ir, executable-specifications-and-lean-proof-envelope | source | source note available; raw/cache text not published | External planning-modeling source for domain/problem separation, action syntax, comparable benchmark notations, and planner-interface discipline. |
ext_shop2_2003 |
SHOP2: An HTN Planning System | external_literature |
planning_htn |
cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); prototype-roadmap (Prototype Roadmap) | planning-as-a-control-layer, intent-to-execution-contracts, cognitive-compilation-and-semantic-ir, prototype-roadmap | source | source note available; raw/cache text not published | External HTN planning source for ordered task decomposition, method selection, temporal/metric planning, and competition-result boundaries. |
ext_integrated_tamp_2020 |
Integrated Task and Motion Planning | external_literature |
planning_task_motion |
Unassigned in current structure | planning-as-a-control-layer, runtime-adapters-tool-permissions-and-human-approval, integrated-reference-architecture | source | source note available; raw/cache text not published | External task-and-motion-planning survey source for discrete task planning, continuous motion planning, black-box subproblem interfaces, and integration-strategy vocabulary. |
ext_behavior_trees_robotics_ai_2017 |
Behavior Trees in Robotics and AI: An Introduction | external_literature |
planning_behavior_trees |
Unassigned in current structure | planning-as-a-control-layer, runtime-adapters-tool-permissions-and-human-approval, integrated-reference-architecture | source | source note available; raw/cache text not published | External behavior-tree source for modular, reactive task switching, robustness/safety analysis vocabulary, planning integration, and stochastic behavior-tree outcome accounting. |
ext_three_states_plan_fear_2006 |
Three States and a Plan: The A.I. of F.E.A.R. | external_literature |
planning_goap |
Unassigned in current structure | planning-as-a-control-layer, runtime-adapters-tool-permissions-and-human-approval, routing-heads-and-specialist-cores | source | source note available; raw/cache text not published | External game-AI planning source for Goal Oriented Action Planning in real-time action games, practical planner constraints, autonomous planning characters, and squad-behavior composition. |
ext_autogen_2023 |
AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation | external_literature |
planning_agent_orchestration |
Unassigned in current structure | planning-as-a-control-layer, labor-os-and-typed-jobs, runtime-adapters-tool-permissions-and-human-approval, artifact-graphs-audit-logs-and-replay | source | source note available; raw/cache text not published | External multi-agent orchestration source for conversable agents, tool/human/LLM operating modes, programmable conversation patterns, and application-level multi-agent workflow boundaries. |
ext_rag_2020 |
Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks | external_literature |
memory_context |
virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates) | virtual-context-abi, context-transactions-snapshots-mounts-and-taint, verification-bandwidth-and-context-adequacy | source | source note available; raw/cache text not published | External retrieval/context source for combining parametric model memory with explicit non-parametric retrieval and provenance-oriented knowledge access. |
ext_lost_in_middle_2023 |
Lost in the Middle: How Language Models Use Long Contexts | external_literature |
memory_context |
virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates) | virtual-context-abi, verification-bandwidth-and-context-adequacy, context-transactions-snapshots-mounts-and-taint | source | source note available; raw/cache text not published | External context-evaluation source for position sensitivity and degraded use of relevant information in the middle of long contexts. |
ext_memgpt_2023 |
MemGPT: Towards LLMs as Operating Systems | external_literature |
memory_context_management |
virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure) | virtual-context-abi, context-transactions-snapshots-mounts-and-taint, verification-bandwidth-and-context-adequacy, procedural-memory-and-cognitive-loop-closure | source | source note available; raw/cache text not published | External memory/context-management source for virtual context management, memory tiers, OS-inspired control flow, and long-running conversation or document-analysis limits. |
ext_longbench_2023 |
LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding | external_literature |
long_context_evaluation |
virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy) | virtual-context-abi, verification-bandwidth-and-context-adequacy, context-transactions-snapshots-mounts-and-taint, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | External long-context benchmark source for multitask long-context understanding, bilingual coverage, retrieval/compression boundaries, and automatic evaluation limits. |
ext_ruler_2024 |
RULER: What’s the Real Context Size of Your Long-Context Language Models? | external_literature |
long_context_evaluation |
virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy) | verification-bandwidth-and-context-adequacy, virtual-context-abi, context-transactions-snapshots-mounts-and-taint, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | External long-context evaluation source for stress-testing context-size claims beyond vanilla needle-in-a-haystack retrieval, including multi-needle, tracing, and aggregation tasks. |
ext_alce_2023 |
Enabling Large Language Models to Generate Text with Citations | external_literature |
retrieval_citation_evaluation |
virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision) | virtual-context-abi, verification-bandwidth-and-context-adequacy, claim-ledgers-and-belief-revision, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | External citation-evaluation source for retrieval-backed answer generation, citation quality metrics, factual correctness, and evidence-support gaps in generated text. |
ext_self_rag_2023 |
Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection | external_literature |
retrieval_reflection |
virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision) | virtual-context-abi, verification-bandwidth-and-context-adequacy, claim-ledgers-and-belief-revision | source | source note available; raw/cache text not published | External retrieval/reflection source for adaptive retrieval, generated critique/reflection tokens, passage relevance, factuality, and citation accuracy boundaries. |
ext_agm_belief_revision_1985 |
On the Logic of Theory Change: Partial Meet Contraction and Revision Functions | external_literature |
belief_revision |
claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision) | claim-ledgers-and-belief-revision | source | source note available; raw/cache text not published | External formal-epistemology comparator for contraction, revision, and AGM-style rational belief change; useful for positioning claim-ledger revision without treating the ASI ledger as an implemented belief-revision engine. |
ext_truth_maintenance_system_1979 |
A Truth Maintenance System | external_literature |
truth_maintenance |
claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision) | claim-ledgers-and-belief-revision | source | source note available; raw/cache text not published | External truth-maintenance comparator for maintaining reasons and justifications for program beliefs; useful for positioning claim ledgers as support-state and revision-history infrastructure, not as implemented truth maintenance. |
ext_assumption_based_tms_1986 |
An Assumption-Based TMS | external_literature |
truth_maintenance |
claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision) | claim-ledgers-and-belief-revision | source | source note available; raw/cache text not published | External assumption-based truth-maintenance comparator for assumption sets, inconsistent information, and context-switching boundaries; useful for distinguishing claim-ledger surface synchronization from implemented ATMS reasoning. |
ext_longllmlingua_2023 |
LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression | external_literature |
context_compression |
virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy) | virtual-context-abi, context-transactions-snapshots-mounts-and-taint, verification-bandwidth-and-context-adequacy, the-efficient-asi-hypothesis, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | External prompt-compression source for long-context cost, latency, position bias, key-information density, and compression/evaluation boundaries. |
ext_proof_carrying_code_1997 |
Proof-Carrying Code | external_literature |
formal_methods |
spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review) | evidence-states-and-claim-discipline, executable-specifications-and-lean-proof-envelope, spinoza-verification-and-proof-carrying-claims, runtime-adapters-tool-permissions-and-human-approval, artifact-graphs-audit-logs-and-replay | source | source note available; raw/cache text not published | External formal-methods source for pairing executable code with machine-checkable evidence that a host can verify against a safety policy. |
ext_tla_plus_home_docs |
My TLA+ Home Page | external_literature |
formal_methods |
Unassigned in current structure | executable-specifications-and-lean-proof-envelope, planning-as-a-control-layer, intent-to-execution-contracts, readiness-gates-residual-escrow-and-quarantine, integrated-reference-architecture | source | source note available; raw/cache text not published | External formal-methods documentation source for TLA+ as a high-level language for modeling programs and systems, especially concurrent and distributed systems. |
ext_lean4_theorem_proving |
Theorem Proving in Lean 4 | external_literature |
formal_methods_proof_assistant |
spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | executable-specifications-and-lean-proof-envelope, circle-calculus-and-proof-carrying-ai-contracts, spinoza-verification-and-proof-carrying-claims, open-research-agenda-and-bibliography-plan | source | source note available; raw/cache text not published | Official Lean theorem-proving text for dependent type theory, propositions, proofs, tactics, inductive types, structures, records, and axioms/computation boundaries. |
ext_autoformalization_llms_2022 |
Autoformalization with Large Language Models | external_literature |
autoformalization |
spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review) | spinoza-verification-and-proof-carrying-claims, executable-specifications-and-lean-proof-envelope, open-research-agenda-and-bibliography-plan | source | source note available; raw/cache text not published | External autoformalization comparator for translating natural-language mathematics into formal specifications and proofs, useful for positioning interpretation-mapping and semantic-adequacy risks in proof-carrying claims. |
ext_ai_safety_debate_2018 |
AI safety via debate | external_literature |
adversarial_review |
scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review) | scalable-oversight-and-adversarial-ai-control, spinoza-verification-and-proof-carrying-claims, policy-optimization-and-learning-from-feedback, open-research-agenda-and-bibliography-plan | source | source note available; raw/cache text not published | External debate comparator for using adversarial agents and a human judge to surface true/useful information when direct human judgment is difficult; useful for positioning tribunal review without treating debate as locally implemented or validated. |
ext_llm_as_judge_mt_bench_2023 |
Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena | external_literature |
model_evaluation |
spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review) | spinoza-verification-and-proof-carrying-claims, benchmark-ratchets-and-anti-goodhart-evidence, policy-optimization-and-learning-from-feedback | source | source note available; raw/cache text not published | External LLM-as-judge comparator for model-graded evaluation, human-preference agreement, and judge bias limits such as position, verbosity, self-enhancement, and reasoning constraints. |
ext_dafny_2010 |
Dafny: An Automatic Program Verifier For Functional Correctness | external_literature |
formal_methods_program_verification |
prototype-roadmap (Prototype Roadmap) | executable-specifications-and-lean-proof-envelope, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval, prototype-roadmap | source | source note available; raw/cache text not published | External program-verification source for specification-oriented programming, functional-correctness verification, SMT-backed automation, and contract/verifier boundaries. |
ext_reluplex_2017 |
Reluplex: An Efficient SMT Solver for Verifying Deep Neural Networks | external_literature |
ai_formal_verification |
adversarial-machine-learning-and-model-attack-surface (Adversarial Machine Learning and the Model Attack Surface); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); resource-economics-and-token-budgets (Resource Economics and Token Budgets); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope) | verification-bandwidth-and-context-adequacy, executable-specifications-and-lean-proof-envelope, benchmark-ratchets-and-anti-goodhart-evidence, resource-economics-and-token-budgets, adversarial-machine-learning-and-model-attack-surface | source | source note available; raw/cache text not published | External AI formal-verification source for SMT-style verification of ReLU neural networks, counterexamples, safety-critical properties, and ACAS Xu evaluation boundaries. |
ext_black_box_simplex_2021 |
The Black-Box Simplex Architecture for Runtime Assurance of Autonomous CPS | external_literature |
formal_runtime_assurance |
Unassigned in current structure | runtime-adapters-tool-permissions-and-human-approval, readiness-gates-residual-escrow-and-quarantine, executable-specifications-and-lean-proof-envelope, integrated-reference-architecture | source | source note available; raw/cache text not published | External runtime-assurance source for switching control authority from advanced controllers to backup safety-preserving behavior under runtime checks. |
ext_copilot_runtime_monitor_2010 |
Copilot: A Hard Real-Time Runtime Monitor | external_literature |
runtime_monitoring |
prototype-roadmap (Prototype Roadmap) | runtime-adapters-tool-permissions-and-human-approval, executable-specifications-and-lean-proof-envelope, readiness-gates-residual-escrow-and-quarantine, prototype-roadmap | source | source note available; raw/cache text not published | External runtime-monitoring source for a stream-based dataflow language/compiler generating constant-time, constant-space C monitors for hard real-time programs. |
ext_cap_theorem_gilbert_lynch_2002 |
Brewer’s Conjecture and the Feasibility of Consistent, Available, Partition-Tolerant Web Services | external_literature |
distributed_systems_consistency |
context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) | personal-compute-hives-and-federated-edge-intelligence, runtime-adapters-tool-permissions-and-human-approval | source | source note available; raw/cache text not published | External distributed-systems source for CAP-style consistency, availability, partition-tolerance, and safety/liveness trade-off vocabulary used to bound partitioned authority, stale grants, and revocation-delay claims without claiming deployed governance consistency. |
ext_prism_model_checker_2002 |
PRISM: Probabilistic Symbolic Model Checker | external_literature |
probabilistic_model_checking |
Unassigned in current structure | executable-specifications-and-lean-proof-envelope, readiness-gates-residual-escrow-and-quarantine, benchmark-ratchets-and-anti-goodhart-evidence, integrated-reference-architecture | source | source note available; raw/cache text not published | External probabilistic model-checking source for symbolic model checking of probabilistic systems, model-checker tooling, and deployment-facing property-analysis vocabulary. |
ext_sparse_moe_2017 |
Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer | external_literature |
routing_modular_intelligence |
Unassigned in current structure | routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis | source | source note available; raw/cache text not published | External MoE/routing source for sparsely-gated expert layers, conditional computation, capacity expansion, load balancing, and routing overhead boundaries. |
ext_gshard_2020 |
GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding | external_literature |
routing_modular_intelligence |
Unassigned in current structure | routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis, integrated-reference-architecture | source | source note available; raw/cache text not published | External MoE/systems source for conditional computation plus automatic sharding, routing, large sparse models, and distributed training constraints. |
ext_switch_transformer_2021 |
Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity | external_literature |
routing_modular_intelligence |
Unassigned in current structure | routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | External MoE/routing source for simplified expert routing, sparse activation, communication/training-stability constraints, and speed/scale claims that require reproduction before local evidence use. |
ext_expert_choice_routing_2022 |
Mixture-of-Experts with Expert Choice Routing | external_literature |
routing_modular_intelligence |
Unassigned in current structure | routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis, readiness-gates-residual-escrow-and-quarantine | source | source note available; raw/cache text not published | External MoE routing source for expert-choice routing, token/expert assignment direction, load-balancing pressure, expert capacity, and convergence/performance claims requiring reproduction before local evidence use. |
ext_mixtral_2024 |
Mixtral of Experts | external_literature |
routing_modular_intelligence |
Unassigned in current structure | routing-heads-and-specialist-cores, fast-generation-architectures, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | External sparse LLM source for token-level expert routing, active-parameter accounting, open MoE model release boundaries, and benchmark claims requiring reproduction before local evidence use. |
ext_moe_llm_survey_2024 |
A Survey on Mixture of Experts in Large Language Models | external_literature |
routing_modular_intelligence |
open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis, open-research-agenda-and-bibliography-plan | source | source note available; raw/cache text not published | External MoE survey source for LLM MoE taxonomy, algorithmic and systemic design issues, implementations, evaluation patterns, and open research directions. |
ext_frugalgpt_2023 |
FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance | external_literature |
task_routing |
Unassigned in current structure | routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis | source | source note available; raw/cache text not published | External task-routing source for prompt adaptation, model approximation, LLM cascades, cost/performance tradeoffs, and query-specific model selection. |
ext_hybrid_llm_2024 |
Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing | external_literature |
cost_quality_routing |
Unassigned in current structure | routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis, readiness-gates-residual-escrow-and-quarantine, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | External query-routing source for predicted query difficulty, small/large model routing, dynamic quality-cost tradeoffs, and quality-preserving large-model-call reduction. |
ext_routellm_2024 |
RouteLLM: Learning to Route LLMs with Preference Data | external_literature |
router_learning |
Unassigned in current structure | routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis, policy-optimization-and-learning-from-feedback | source | source note available; raw/cache text not published | External learned-router source for routing between stronger and weaker LLMs using preference data, cost-quality tradeoffs, and transfer to changed model pairs. |
ext_deep_compression_2015 |
Deep Compression: Compressing Deep Neural Networks with Pruning, Trained Quantization and Huffman Coding | external_literature |
compression_representation |
Unassigned in current structure | compact-generative-systems-and-residual-honesty, rankfold-neuralfold-and-artifact-compression, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | External compression source for pruning, trained quantization, coding, memory-footprint reduction, and speed/energy claims that require reproduction before local evidence use. |
ext_lora_2021 |
LoRA: Low-Rank Adaptation of Large Language Models | external_literature |
compression_representation |
Unassigned in current structure | rankfold-neuralfold-and-artifact-compression, compact-generative-systems-and-residual-honesty, resource-economics-and-token-budgets, policy-optimization-and-learning-from-feedback, coilra-multicoil-rope-and-cyclic-mixers | source | source note available; raw/cache text not published | External low-rank adaptation source for parameter-efficient updates, rank-decomposition adapters, memory reduction, and adaptation-boundary vocabulary. |
ext_knowledge_distillation_2015 |
Distilling the Knowledge in a Neural Network | external_literature |
compression_representation |
Unassigned in current structure | compact-generative-systems-and-residual-honesty, rankfold-neuralfold-and-artifact-compression, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | External compression source for teacher/student distillation, soft-target transfer, ensemble compression, and knowledge-transfer claims requiring local reproduction before evidence use. |
ext_gptq_2022 |
GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers | external_literature |
compression_quantization |
Unassigned in current structure | rankfold-neuralfold-and-artifact-compression, compact-generative-systems-and-residual-honesty, resource-economics-and-token-budgets, fast-generation-architectures | source | source note available; raw/cache text not published | External quantization source for post-training compression of large generative transformers, one-shot weight quantization, memory reduction, and accuracy/speed tradeoff boundaries. |
ext_qlora_2023 |
QLoRA: Efficient Finetuning of Quantized LLMs | external_literature |
compression_quantized_adaptation |
prototype-roadmap (Prototype Roadmap) | rankfold-neuralfold-and-artifact-compression, resource-economics-and-token-budgets, policy-optimization-and-learning-from-feedback, prototype-roadmap | source | source note available; raw/cache text not published | External quantized-adaptation source for finetuning quantized LLMs with low-rank adapters, memory-efficient training, and benchmark claims requiring reproduction before local evidence use. |
ext_dreamcoder_2020 |
DreamCoder: Growing generalizable, interpretable knowledge with wake-sleep Bayesian program learning | external_literature |
program_synthesis_representation |
cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | cognitive-compilation-and-semantic-ir, procedural-memory-and-cognitive-loop-closure, compact-generative-systems-and-residual-honesty, rankfold-neuralfold-and-artifact-compression, mathematical-and-search-substrates, replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | External program-synthesis source for wake-sleep library learning, reusable abstractions, interpretable learned programs, and compression-through-abstraction vocabulary. |
ext_llvm_langref_docs |
LLVM Language Reference Manual | external_literature |
compiler_ir |
cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR) | cognitive-compilation-and-semantic-ir, executable-specifications-and-lean-proof-envelope | source | source note available; raw/cache text not published | Official LLVM Language Reference comparator for SSA-based intermediate representation, equivalent in-memory/bitcode/human-readable forms, well-formedness, verifier passes, and optimization/analysis vocabulary. |
ext_mlir_2020 |
MLIR: A Compiler Infrastructure for the End of Moore’s Law | external_literature |
multi_level_compiler_ir |
cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR) | cognitive-compilation-and-semantic-ir, resource-economics-and-token-budgets, mathematical-and-search-substrates | source | source note available; raw/cache text not published | External multi-level compiler-IR comparator for reusable and extensible compiler infrastructure, dialects, progressive lowering, verifiers, modular passes, and heterogeneous target support. |
ext_translation_validation_1998 |
Translation Validation | external_literature |
translation_validation |
cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR) | cognitive-compilation-and-semantic-ir, executable-specifications-and-lean-proof-envelope | source | source note available; raw/cache text not published | External translation-validation comparator for checking each compiler/code-generator run after translation, using a common semantic framework, refinement relation, and simulation-based proof method. |
ext_toolformer_2023 |
Toolformer: Language Models Can Teach Themselves to Use Tools | external_literature |
learned_tool_use |
procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure) | procedural-memory-and-cognitive-loop-closure, runtime-adapters-tool-permissions-and-human-approval, policy-optimization-and-learning-from-feedback | source | source note available; raw/cache text not published | External learned-tool-use source for self-supervised API-call insertion, tool selection, argument construction, and result incorporation without claiming ASI Stack tool-use reproduction. |
ext_voyager_2023 |
Voyager: An Open-Ended Embodied Agent with Large Language Models | external_literature |
lifelong_skill_learning |
open-ended-improvement-engines (Open-Ended Improvement Engines); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure) | procedural-memory-and-cognitive-loop-closure, routing-heads-and-specialist-cores, benchmark-ratchets-and-anti-goodhart-evidence, open-ended-improvement-engines | source | source note available; raw/cache text not published | External lifelong-agent source for automatic curriculum, executable-code skill libraries, iterative environment-feedback prompting, self-verification, and skill-library transfer in Minecraft. |
ext_information_bottleneck_2000 |
The information bottleneck method | external_literature |
representation_compression |
learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) | compact-generative-systems-and-residual-honesty, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | External representation-compression source for relevance-preserving compression, bottleneck variables, mutual-information tradeoffs, and compression/utility separation. |
ext_mdl_tutorial_2004 |
A tutorial introduction to the minimum description length principle | external_literature |
description_length_residuals |
learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) | compact-generative-systems-and-residual-honesty, rankfold-neuralfold-and-artifact-compression, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | External description-length source for model/data tradeoffs, compression as inductive discipline, and residual/error-accounting vocabulary. |
ext_weakness_generalization_2023 |
The Optimal Choice of Hypothesis Is the Weakest, Not the Shortest | external_literature |
learning_theory_inductive_bias |
learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) | learning-theory-generalization-and-scaling-science | source | source note available; raw/cache text not published | Bennett’s finite enactive-cognition formalism separates extension-based hypothesis weakness from description length. Under a uniform distribution over its task space, the paper argues that maximizing weakness is necessary and sufficient for maximizing generalization probability and gives a counterexample to MDL as a universal proxy. Its theorem assumptions and toy 8-bit arithmetic experiments do not establish a general result for neural networks or real task distributions. |
ext_codebleu_2020 |
CodeBLEU: a Method for Automatic Evaluation of Code Synthesis | external_literature |
artifact_utility_metrics |
prototype-roadmap (Prototype Roadmap) | benchmark-ratchets-and-anti-goodhart-evidence, compact-generative-systems-and-residual-honesty, artifact-steward-agents-and-living-project-governance, prototype-roadmap | source | source note available; raw/cache text not published | External code-synthesis evaluation source for combining lexical, syntax, data-flow, and semantic matching into artifact-quality metrics that still require task-specific validation. |
ext_mmlu_2020 |
Measuring Massive Multitask Language Understanding | external_literature |
benchmark_science |
prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | benchmark-ratchets-and-anti-goodhart-evidence, evidence-states-and-claim-discipline, open-research-agenda-and-bibliography-plan, prototype-roadmap | source | source note available; raw/cache text not published | External benchmark source for broad multitask evaluation, task-coverage limits, lopsided performance, uncertainty about wrong answers, and benchmark-saturation pressure. |
ext_bigbench_2022 |
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models | external_literature |
benchmark_science |
Unassigned in current structure | benchmark-ratchets-and-anti-goodhart-evidence, the-efficient-asi-hypothesis, readiness-gates-residual-escrow-and-quarantine, policy-optimization-and-learning-from-feedback | source | source note available; raw/cache text not published | External benchmark source for BIG-bench, broad task coverage, scale effects, calibration, breakthrough behavior, human-rater baselines, and social-bias tradeoffs. |
ext_helm_2022 |
Holistic Evaluation of Language Models | external_literature |
benchmark_science |
living-book-methodology (Living Book Methodology) | benchmark-ratchets-and-anti-goodhart-evidence, evidence-states-and-claim-discipline, readiness-gates-residual-escrow-and-quarantine, living-book-methodology | source | source note available; raw/cache text not published | External benchmark-science source for multi-scenario, multi-metric evaluation, missing-coverage disclosure, raw-prompt transparency, and living benchmark practice. |
ext_gpqa_2023 |
GPQA: A Graduate-Level Google-Proof Q&A Benchmark | external_literature |
benchmark_science |
verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | benchmark-ratchets-and-anti-goodhart-evidence, verification-bandwidth-and-context-adequacy, evidence-states-and-claim-discipline, open-research-agenda-and-bibliography-plan | source | source note available; raw/cache text not published | External benchmark source for expert-written hard questions, Google-proof validation, scalable oversight pressure, and the gap between skilled non-expert validation and expert competence. |
ext_swe_bench_2023 |
SWE-bench: Can Language Models Resolve Real-World GitHub Issues? | external_literature |
benchmark_science |
prototype-roadmap (Prototype Roadmap) | benchmark-ratchets-and-anti-goodhart-evidence, artifact-graphs-audit-logs-and-replay, labor-os-and-typed-jobs, prototype-roadmap | source | source note available; raw/cache text not published | External benchmark source for real-world software-engineering issue resolution, repository-scale context, executable environments, patch evaluation, and capability boundaries. |
ext_swe_rebench_v2_2026 |
SWE-rebench V2: Language-Agnostic SWE Task Collection at Scale | external_literature |
natural_software_task_construction_and_evaluation |
artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap) | benchmark-ratchets-and-anti-goodhart-evidence, artifact-graphs-audit-logs-and-replay, integrated-reference-architecture, prototype-roadmap, readiness-gates-residual-escrow-and-quarantine | source | source note available; raw/cache text not published | Primary 2026 natural-task substrate for multilingual repository changes, interactive setup, containerized full-suite execution, separated solution/test patches, task diagnostics, and explicit environment pathologies. It does not establish local task validity, gold execution, model competence, governance benefit, safety, transfer, or SOTA. |
ext_livebench_2024 |
LiveBench: A Challenging, Contamination-Limited LLM Benchmark | external_literature |
benchmark_science |
living-book-methodology (Living Book Methodology); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | benchmark-ratchets-and-anti-goodhart-evidence, evidence-states-and-claim-discipline, living-book-methodology, open-research-agenda-and-bibliography-plan | source | source note available; raw/cache text not published | External benchmark source for contamination-limited evaluation, frequently updated questions, objective ground-truth scoring, and monthly benchmark evolution. |
ext_dynabench_2021 |
Dynabench: Rethinking Benchmarking in NLP | external_literature |
dynamic_benchmarking |
Unassigned in current structure | benchmark-ratchets-and-anti-goodhart-evidence, readiness-gates-residual-escrow-and-quarantine, policy-optimization-and-learning-from-feedback, artifact-steward-agents-and-living-project-governance | source | source note available; raw/cache text not published | External dynamic-benchmarking source for human-and-model-in-the-loop data collection, adversarial benchmark evolution, and stale static benchmark pressure. |
ext_checklist_2020 |
Beyond Accuracy: Behavioral Testing of NLP models with CheckList | external_literature |
behavioral_evaluation |
verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); prototype-roadmap (Prototype Roadmap) | benchmark-ratchets-and-anti-goodhart-evidence, verification-bandwidth-and-context-adequacy, claim-ledgers-and-belief-revision, prototype-roadmap | source | source note available; raw/cache text not published | External behavioral-testing source for capability matrices, minimum functionality tests, invariance tests, directional expectation tests, and failure-discovery beyond aggregate accuracy. |
ext_benchmark_contamination_2023 |
Investigating Data Contamination in Modern Benchmarks for Large Language Models | external_literature |
benchmark_contamination |
living-book-methodology (Living Book Methodology) | benchmark-ratchets-and-anti-goodhart-evidence, evidence-states-and-claim-discipline, readiness-gates-residual-escrow-and-quarantine, living-book-methodology | source | source note available; raw/cache text not published | External benchmark-contamination source for detecting training/test overlap pressure, benchmark-leakage risk, and score interpretation limits in modern LLM evaluations. |
ext_goodhart_variants_2018 |
Categorizing Variants of Goodhart’s Law | external_literature |
goodhart_failure_taxonomy |
failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence) | failure-modes-of-ungoverned-intelligence, benchmark-ratchets-and-anti-goodhart-evidence, policy-optimization-and-learning-from-feedback, artifact-steward-agents-and-living-project-governance, evidence-states-and-claim-discipline | source | source note available; raw/cache text not published | External Goodhart-taxonomy source for regressive, extremal, causal, and adversarial metric failures that benchmark ratchets and policy updates must treat as distinct risks. |
ext_speculative_decoding_2022 |
Fast Inference from Transformers via Speculative Decoding | external_literature |
fast_generation |
fast-generation-architectures (Fast Generation Architectures) | fast-generation, the-efficient-asi-hypothesis | source | source note available; raw/cache text not published | Primary external paper for speculative decoding: a draft model proposes multiple tokens and a target model verifies them, giving an exact-distribution acceleration path under its assumptions. |
ext_multi_token_prediction_2024 |
Better & Faster Large Language Models via Multi-token Prediction | external_literature |
fast_generation |
fast-generation-architectures (Fast Generation Architectures) | fast-generation, the-efficient-asi-hypothesis | source | source note available; raw/cache text not published | Primary external paper for multi-token prediction as an auxiliary training objective and inference-time multi-token proposal mechanism. |
ext_medusa_2024 |
Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads | external_literature |
fast_generation |
fast-generation-architectures (Fast Generation Architectures) | fast-generation, the-efficient-asi-hypothesis | source | source note available; raw/cache text not published | Primary external paper for adding multiple decoding heads to an LLM and verifying tree-structured candidate continuations in parallel. |
ext_eagle_2024 |
EAGLE: Speculative Sampling Requires Rethinking Feature Uncertainty | external_literature |
fast_generation |
fast-generation-architectures (Fast Generation Architectures) | fast-generation, the-efficient-asi-hypothesis | source | source note available; raw/cache text not published | Primary external paper for feature-level speculative drafting and target-model verification as an acceleration mechanism. |
ext_lookahead_decoding_2024 |
Break the Sequential Dependency of LLM Inference Using Lookahead Decoding | external_literature |
fast_generation |
fast-generation-architectures (Fast Generation Architectures) | fast-generation | source | source note available; raw/cache text not published | Primary external paper for lookahead decoding: a parallel exact decoding algorithm that reduces serial decoding steps without an auxiliary draft model. |
ext_layerskip_2024 |
LayerSkip: Enabling Early Exit Inference and Self-Speculative Decoding | external_literature |
fast_generation |
fast-generation-architectures (Fast Generation Architectures) | fast-generation | source | source note available; raw/cache text not published | Primary external paper for early-exit inference and self-speculative decoding where early layers draft and later layers verify. |
ext_pagedattention_vllm_2023 |
Efficient Memory Management for Large Language Model Serving with PagedAttention | external_literature |
fast_generation |
virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation, resource-economics | source | source note available; raw/cache text not published | Primary external paper for vLLM/PagedAttention, which treats KV-cache memory management and serving throughput as a distinct acceleration axis. |
ext_transformer_xl_2019 |
Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context | external_literature |
sequence_memory_recurrence |
Unassigned in current structure | coil-attention-cyclic-memory-and-recurrence-contracts | source | source note available; raw/cache text not published | External recurrent Transformer comparator for segment-level recurrence, relative positional encoding, and long-dependency language modeling; useful as a baseline family for cyclic-memory contracts without implying local reproduction. |
ext_compressive_transformer_2019 |
Compressive Transformers for Long-Range Sequence Modelling | external_literature |
sequence_memory_recurrence |
Unassigned in current structure | coil-attention-cyclic-memory-and-recurrence-contracts | source | source note available; raw/cache text not published | External long-range memory comparator for compressed past memories, memory mechanisms, and long-range sequence benchmarks; useful for positioning cyclic memory against compression-memory baselines. |
ext_roformer_rope_2021 |
RoFormer: Enhanced Transformer with Rotary Position Embedding | external_literature |
position_encoding |
Unassigned in current structure | coilra-multicoil-rope-and-cyclic-mixers | source | source note available; raw/cache text not published | External RoPE comparator for rotary position embedding, relative-position behavior inside self-attention, and position-encoding baselines for cyclic phase or RoPE-style substrates. |
ext_retnet_2023 |
Retentive Network: A Successor to Transformer for Large Language Models | external_literature |
sequence_memory_recurrence |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | coil-attention-cyclic-memory-and-recurrence-contracts, coilra-multicoil-rope-and-cyclic-mixers, replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | External retention/recurrent-sequence comparator for the relationship between recurrence and attention, recurrent/chunkwise computation, and inference-efficiency tradeoffs. |
ext_mamba_2023 |
Mamba: Linear-Time Sequence Modeling with Selective State Spaces | external_literature |
sequence_substrates |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); fast-generation-architectures (Fast Generation Architectures); mathematical-and-search-substrates (Mathematical and Search Substrates) | fast-generation, math-substrates, coilra-multicoil-rope-and-cyclic-mixers, replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary external paper for selective state-space sequence models as a different long-sequence substrate and inference-efficiency axis from decoding tricks. |
ext_llada_2025 |
Large Language Diffusion Models | external_literature |
diffusion_language_models |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); fast-generation-architectures (Fast Generation Architectures) | fast-generation-architectures, replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary external paper for LLaDA, a large masked-diffusion language model trained with pretraining and supervised fine-tuning rather than left-to-right autoregression. |
ext_scaling_dllms_2026 |
Scaling Beyond Masked Diffusion Language Models | external_literature |
diffusion_language_models |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); fast-generation-architectures (Fast Generation Architectures) | fast-generation-architectures, replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary external paper for comparing diffusion language-model families by speed-quality tradeoffs rather than perplexity alone. |
ext_trpo_2015 |
Trust Region Policy Optimization | external_literature |
policy_optimization |
policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | governed-deliberation-and-test-time-scaling | source | source note available; raw/cache text not published | Primary external source for trust-region policy-gradient updates and bounded update-size discipline. |
ext_ppo_2017 |
Proximal Policy Optimization Algorithms | external_literature |
policy_optimization |
policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | governed-deliberation-and-test-time-scaling | source | source note available; raw/cache text not published | Primary external source for PPO-style online policy-gradient updates and proximal surrogate objectives. |
ext_remax_2023 |
ReMax: A Simple, Effective, and Efficient Reinforcement Learning Method for Aligning Large Language Models | external_literature |
policy_optimization |
policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | source | source note available; raw/cache text not published | Primary external source for simpler RLHF-oriented policy-gradient updates relative to PPO-style machinery. | |
ext_goal_oriented_requirements_engineering_2001 |
Goal-Oriented Requirements Engineering: A Guided Tour | external_literature |
requirements_engineering |
human-intent-as-a-formal-input (Human Intent as a Formal Input) | human-intent-as-a-formal-input | source | source note available; raw/cache text not published | External requirements-engineering comparator for turning stakeholder goals, constraints, refinements, and responsibilities into explicit requirements before system design or execution. |
ext_cooperative_inverse_rl_2016 |
Cooperative Inverse Reinforcement Learning | external_literature |
human_intent_alignment |
human-intent-as-a-formal-input (Human Intent as a Formal Input); governed-objective-formation-value-learning-and-goal-integrity (Governed Objective Formation, Value Learning, and Goal Integrity) | human-intent-as-a-formal-input | source | source note available; raw/cache text not published | External cooperative AI comparator for formalizing value alignment as uncertainty about the human reward function in a cooperative partial-information setting. |
ext_deep_rl_human_preferences_2017 |
Deep Reinforcement Learning from Human Preferences | external_literature |
human_feedback_learning |
human-intent-as-a-formal-input (Human Intent as a Formal Input) | human-intent-as-a-formal-input, policy-optimization-and-learning-from-feedback | source | source note available; raw/cache text not published | External human-feedback comparator for communicating complex goals through preference comparisons over behavior segments; useful for separating preference signals from explicit intent contracts. |
ext_dpo_2023 |
Direct Preference Optimization: Your Language Model is Secretly a Reward Model | external_literature |
policy_optimization |
policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | source | source note available; raw/cache text not published | Primary external source for DPO-style offline preference optimization without a separate online RL loop. | |
ext_ipo_preference_2023 |
A General Theoretical Paradigm to Understand Learning from Human Preferences | external_literature |
policy_optimization |
policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | source | source note available; raw/cache text not published | External source for theoretical comparison of preference-learning objectives, including IPO/DPO-style framing. | |
ext_orpo_2024 |
ORPO: Monolithic Preference Optimization without Reference Model | external_literature |
policy_optimization |
policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | source | source note available; raw/cache text not published | External source for reference-model-free monolithic preference optimization. | |
ext_kto_2024 |
KTO: Model Alignment as Prospect Theoretic Optimization | external_literature |
policy_optimization |
policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | source | source note available; raw/cache text not published | External source for prospect-theoretic model alignment and human-aware loss framing. | |
ext_simpo_2024 |
SimPO: Simple Preference Optimization with a Reference-Free Reward | external_literature |
policy_optimization |
policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | source | source note available; raw/cache text not published | External source for simple reference-free preference optimization using sequence-level reward framing. | |
ext_reinforce_style_rlhf_2024 |
Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs | external_literature |
policy_optimization |
policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | source | source note available; raw/cache text not published | External source for revisiting simpler REINFORCE-style optimization as an RLHF baseline. | |
ext_deepseek_r1_2025 |
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning | external_literature |
policy_optimization |
governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | source | source note available; raw/cache text not published | External source for reinforcement-learning pressure on reasoning behavior in large language models. | |
ext_dapo_2025 |
DAPO: An Open-Source LLM Reinforcement Learning System at Scale | external_literature |
policy_optimization |
policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | source | source note available; raw/cache text not published | External source for open-source large-scale LLM RL system details and DAPO-style update design. | |
ext_gspo_2025 |
Group Sequence Policy Optimization | external_literature |
policy_optimization |
policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | source | source note available; raw/cache text not published | External source for sequence-level group policy optimization in LLM reinforcement learning. | |
ext_s_grpo_2025 |
S-GRPO: Early Exit via Reinforcement Learning in Reasoning Models | external_literature |
policy_optimization |
governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | source | source note available; raw/cache text not published | External source for early-exit reinforcement learning and overthinking control in reasoning models. | |
ext_longrlvr_2026 |
LongRLVR: Long-Context Reinforcement Learning Requires Verifiable Context Rewards | external_literature |
policy_optimization |
policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | source | source note available; raw/cache text not published | External source for long-context RLVR and verifiable context-grounding rewards. | |
ext_rlhf_limitations_2023 |
Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback | external_literature |
policy_optimization |
policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | source | source note available; raw/cache text not published | External survey source for RLHF limitations, reward hacking, evaluator limits, and complementary safeguards. | |
ext_tailscale_docs_2025 |
What is Tailscale? | external_literature |
personal_compute_hives |
personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) | personal-compute-hives-and-federated-edge-intelligence | source | source note available; raw/cache text not published | Official Tailscale documentation for zero-trust identity networking, tailnets, encrypted point-to-point connections, and cross-network device connectivity. |
ext_kubernetes_overview_docs |
Kubernetes Documentation: Overview | external_literature |
personal_compute_hives |
personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) | personal-compute-hives-and-federated-edge-intelligence | source | source note available; raw/cache text not published | Official Kubernetes overview for containerized workload management, declarative configuration, automation, service discovery, storage orchestration, rollouts, bin packing, and self-healing. |
ext_k3s_docs_2026 |
K3s: Lightweight Kubernetes | external_literature |
personal_compute_hives |
personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) | personal-compute-hives-and-federated-edge-intelligence | source | source note available; raw/cache text not published | Official K3s documentation for lightweight Kubernetes deployment in edge, homelab, IoT, CI, single-board-computer, air-gapped, and embedded settings. |
ext_nomad_docs |
Nomad Documentation | external_literature |
personal_compute_hives |
personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) | personal-compute-hives-and-federated-edge-intelligence | source | source note available; raw/cache text not published | Official Nomad documentation for scheduling and orchestrating containers, non-containerized applications, and batch jobs across on-prem and cloud environments. |
ext_temporal_docs |
Temporal Documentation: What is Temporal? | external_literature |
durable_execution |
labor-os-and-typed-jobs (Labor OS and Typed Jobs) | intent-to-execution-contracts, labor-os-and-typed-jobs | source | source note available; raw/cache text not published | Official Temporal documentation comparator for durable workflow execution, workflow event histories, worker processes, failure recovery, and long-running application-code orchestration. |
ext_airflow_dag_docs |
Apache Airflow Documentation: Dags | external_literature |
workflow_orchestration |
labor-os-and-typed-jobs (Labor OS and Typed Jobs) | intent-to-execution-contracts, labor-os-and-typed-jobs | source | source note available; raw/cache text not published | Official Apache Airflow documentation comparator for DAG-based workflow scheduling, tasks, dependencies, callbacks, retries, and operational workflow metadata. |
ext_bpmn_2_0_2_spec |
Business Process Model and Notation Specification Version 2.0.2 | external_literature |
business_process_modeling |
labor-os-and-typed-jobs (Labor OS and Typed Jobs) | intent-to-execution-contracts, labor-os-and-typed-jobs | source | source note available; raw/cache text not published | OMG BPMN 2.0.2 formal specification comparator for stakeholder-readable business-process diagrams, implementation-independent flow notation, and translation into software process components. |
ext_kubernetes_jobs_docs |
Kubernetes Documentation: Jobs | external_literature |
batch_job_lifecycle |
labor-os-and-typed-jobs (Labor OS and Typed Jobs) | labor-os-and-typed-jobs | source | source note available; raw/cache text not published | Official Kubernetes Jobs documentation comparator for batch job lifecycle, completions, backoff limits, active deadlines, terminal Complete/Failed conditions, and cleanup of finished jobs. |
ext_ray_core_docs_2026 |
What’s Ray Core? | external_literature |
personal_compute_hives |
personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) | personal-compute-hives-and-federated-edge-intelligence | source | source note available; raw/cache text not published | Official Ray Core documentation for distributed task, actor, and object primitives used to build and scale Python applications. |
ext_boinc_home_2026 |
BOINC | external_literature |
personal_compute_hives |
personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) | personal-compute-hives-and-federated-edge-intelligence | source | source note available; raw/cache text not published | Official BOINC site for volunteer computing where user computers download scientific computing jobs and run them in the background. |
ext_syncthing_home |
Syncthing | external_literature |
personal_compute_hives |
personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) | personal-compute-hives-and-federated-edge-intelligence | source | source note available; raw/cache text not published | Official Syncthing site for continuous file synchronization across computers, authenticated devices, encrypted transport, and user-controlled storage location. |
ext_ipfs_docs |
IPFS Documentation and Project Site | external_literature |
personal_compute_hives |
personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) | personal-compute-hives-and-federated-edge-intelligence | source | source note available; raw/cache text not published | Official IPFS project documentation and site for peer-to-peer content addressing, content identifiers, provider discovery, and decentralized retrieval vocabulary. |
ext_akash_docs_2026 |
Akash Network Documentation | external_literature |
personal_compute_hives |
personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) | personal-compute-hives-and-federated-edge-intelligence, artifact-steward-agents-and-living-project-governance | source | source note available; raw/cache text not published | Official Akash documentation for decentralized cloud deployment, provider resources, leases, GPUs, SDKs, node operations, and provider operations. |
ext_golem_docs_2025 |
Golem Developer Resources | external_literature |
personal_compute_hives |
personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) | personal-compute-hives-and-federated-edge-intelligence, artifact-steward-agents-and-living-project-governance | source | source note available; raw/cache text not published | Official Golem developer resources for decentralized computations, task execution, provider selection, result handling, and resource sharing. |
ext_github_webhooks_docs |
Webhook events and payloads | external_literature |
artifact_steward_agents |
artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) | artifact-steward-agents-and-living-project-governance | source | source note available; raw/cache text not published | Official GitHub documentation for repository and organization webhook events, event payloads, delivery headers, event-specific permissions, and payload limits. |
ext_github_self_hosted_runners_docs |
Self-hosted runners | external_literature |
artifact_steward_agents |
personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) | artifact-steward-agents-and-living-project-governance, personal-compute-hives-and-federated-edge-intelligence | source | source note available; raw/cache text not published | Official GitHub Actions documentation for self-hosted runners as user-managed systems that execute workflow jobs on physical, virtual, containerized, on-prem, or cloud machines. |
ext_openzeppelin_governor_docs |
OpenZeppelin Contracts: Governance | external_literature |
artifact_steward_agents |
artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) | artifact-steward-agents-and-living-project-governance | source | source note available; raw/cache text not published | Official OpenZeppelin governance documentation for modular Governor contracts, voting power, quorum, timelocks, proposal settings, and guardian-style extensions. |
ext_open_collective_docs |
Open Collective Documentation | external_literature |
artifact_steward_agents |
artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) | artifact-steward-agents-and-living-project-governance | source | source note available; raw/cache text not published | Official Open Collective documentation for transparent community money management, fiscal hosts, contribution intake, expenses, accounting, and legal-entity delegation through fiscal hosting. |
ext_github_sponsors_docs |
About GitHub Sponsors for open source contributors | external_literature |
artifact_steward_agents |
artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) | artifact-steward-agents-and-living-project-governance | source | source note available; raw/cache text not published | Official GitHub Sponsors documentation for contributor and organization sponsorship eligibility, open-source contribution categories, sponsor profiles, and GitHub-native funding surfaces. |
ext_agentic_workflow_injection_2026 |
Demystifying and Detecting Agentic Workflow Injection Vulnerabilities in GitHub Actions | external_literature |
artifact_steward_agents |
artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) | artifact-steward-agents-and-living-project-governance | source | source note available; raw/cache text not published | External security paper for agentic workflow injection in GitHub Actions when untrusted repository event context reaches LLM agents and downstream workflow logic. |
ext_dao_delegation_fairness_2025 |
Fairness in Token Delegation: Mitigating Voting Power Concentration in DAOs | external_literature |
artifact_steward_agents |
artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) | artifact-steward-agents-and-living-project-governance | source | source note available; raw/cache text not published | External DAO governance paper for voter apathy, voting-power concentration, delegation misalignment, and delegate-ranking bias. |
ext_model_cards_2019 |
Model Cards for Model Reporting | external_literature |
model_reporting |
project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference) | evidence-states-and-claim-discipline, project-theseus-as-report-first-implementation-reference | source | source note available; raw/cache text not published | External reporting comparator for structured model documentation, intended-use boundaries, evaluation disclosures, ethical considerations, and model-report artifacts. |
ext_datasheets_datasets_2021 |
Datasheets for Datasets | external_literature |
dataset_documentation |
project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference) | evidence-states-and-claim-discipline, project-theseus-as-report-first-implementation-reference | source | source note available; raw/cache text not published | External documentation comparator for dataset motivation, composition, collection, preprocessing, uses, distribution, maintenance, and accountability questions. |
ext_factsheets_ai_services_2019 |
FactSheets: Increasing Trust in AI Services through Supplier’s Declarations of Conformity | external_literature |
ai_service_fact_sheets |
project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference) | project-theseus-as-report-first-implementation-reference | source | source note available; raw/cache text not published | External AI-service documentation comparator for supplier declarations, service properties, trust-relevant facts, and standardized reporting boundaries. |
ext_ml_reproducibility_program_2021 |
Improving Reproducibility in Machine Learning Research | external_literature |
ml_reproducibility_reporting |
project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference) | evidence-states-and-claim-discipline, project-theseus-as-report-first-implementation-reference | source | source note available; raw/cache text not published | External reproducibility-program comparator for checklists, code submission, reproducibility reports, and community review mechanisms in machine-learning research. |
ext_transformer_circuits_2021 |
A Mathematical Framework for Transformer Circuits | external_literature |
mechanistic_interpretability |
white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) | evidence-states-and-claim-discipline, white-box-evidence-interpretability-and-activation-governance | source | source note available; raw/cache text not published | External mechanistic-interpretability comparator for treating internal circuit analyses as scoped white-box evidence that still needs model, layer, behavior, causal, and limitation boundaries. |
ext_monosemanticity_2023 |
Towards Monosemanticity: Decomposing Language Models With Dictionary Learning | external_literature |
mechanistic_interpretability |
white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) | evidence-states-and-claim-discipline, white-box-evidence-interpretability-and-activation-governance | source | source note available; raw/cache text not published | External mechanistic-interpretability comparator for sparse-autoencoder feature decomposition and the boundary between discovered features, feature-level evidence, and broader model-behavior claims. |
ext_literate_programming_1984 |
Literate Programming | external_literature |
literate_programming |
living-book-methodology (Living Book Methodology) | living-book-methodology | source | source note available; raw/cache text not published | External literate-programming source for arranging programs and explanation around human comprehension, woven documentation, and tangled executable artifacts. |
ext_jupyter_book_docs |
Jupyter Book Documentation | external_literature |
executable_books |
living-book-methodology (Living Book Methodology) | living-book-methodology | source | source note available; raw/cache text not published | Official Jupyter Book documentation comparator for authoring books from notebooks or Markdown, executing code, cross-referencing content, and publishing computational books to the web. |
ext_quarto_books_docs |
Quarto Books Documentation | external_literature |
technical_publishing |
living-book-methodology (Living Book Methodology) | living-book-methodology | source | source note available; raw/cache text not published | Official Quarto Books documentation comparator for multi-chapter manuscripts, HTML/PDF/Word/EPUB outputs, search, cross references, and book-style website publishing. |
ext_argo_rollouts_docs |
Argo Rollouts Documentation: Kubernetes Progressive Delivery Controller | external_literature |
progressive_delivery_rollback |
capability-replacement-and-rollback (Capability Replacement and Rollback) | capability-replacement-and-rollback | source | source note available; raw/cache text not published | External progressive-delivery comparator for blue-green rollout, canary rollout, metric analysis, automated promotion, and automated rollback vocabulary. |
ext_feature_toggles_fowler |
Feature Toggles (aka Feature Flags) | external_literature |
feature_flag_release_control |
capability-replacement-and-rollback (Capability Replacement and Rollback) | capability-replacement-and-rollback | source | source note available; raw/cache text not published | External feature-flag comparator for controlled exposure, canary releasing, release toggles, experiment toggles, ops toggles, permissioning toggles, and validation complexity. |
ext_google_cloud_mlops_cd |
MLOps: Continuous Delivery and Automation Pipelines in Machine Learning | external_literature |
mlops_continuous_delivery |
capability-replacement-and-rollback (Capability Replacement and Rollback) | capability-replacement-and-rollback | source | source note available; raw/cache text not published | External MLOps comparator for CI/CD/CT, data/model validation, model deployment, monitoring, rollback triggers, and model-regression concerns. |
ext_kubernetes_deployments_docs |
Kubernetes Documentation: Deployments | external_literature |
deployment_rollout_rollback |
capability-replacement-and-rollback (Capability Replacement and Rollback) | capability-replacement-and-rollback | source | source note available; raw/cache text not published | External deployment-controller comparator for rollout status, rollout history, revision records, and rollback to a prior stable Deployment revision. |
ext_drexler_cais_2019 |
Reframing Superintelligence: Comprehensive AI Services as General Intelligence | external_literature |
ai_services_r_and_d_automation |
asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); integrated-reference-architecture (Integrated Reference Architecture) | asi-is-a-stack-not-a-model, constitutional-alignment-substrate, recursive-self-improvement-boundaries, integrated-reference-architecture | source | source note available; raw/cache text not published | Primary CAIS technical-report comparator for service-centered general intelligence, R&D automation, structured AI development, and the distinction between component affordances and a complete governance architecture. |
ext_mrkl_systems_2022 |
MRKL Systems: A Modular, Neuro-Symbolic Architecture That Combines Large Language Models, External Knowledge Sources and Discrete Reasoning | external_literature |
neuro_symbolic_modular_architecture |
Unassigned in current structure | asi-is-a-stack-not-a-model | source | source note available; raw/cache text not published | External modular-neuro-symbolic architecture comparator for combining language models with expert modules, external knowledge sources, and routing rather than treating the model as the whole system. |
ext_llm_agents_survey_2023 |
A Survey on Large Language Model based Autonomous Agents | external_literature |
llm_agent_architecture |
Unassigned in current structure | asi-is-a-stack-not-a-model | source | source note available; raw/cache text not published | External LLM-agent architecture comparator for agent profiles, memory, planning, and action modules around a language model, useful for positioning the stack frame against agent-system decompositions. |
ext_standard_model_mind_2017 |
A Standard Model of the Mind: Toward a Common Computational Framework Across Artificial Intelligence, Cognitive Science, Neuroscience, and Robotics | external_literature |
cognitive_architecture |
Unassigned in current structure | asi-is-a-stack-not-a-model | source | source note available; raw/cache text not published | External cognitive-architecture comparator for treating intelligent behavior as an integrated architecture spanning memory, learning, perception/action, procedural control, and deliberation. |
ext_subsumption_architecture_1986 |
A Robust Layered Control System for a Mobile Robot | external_literature |
layered_robot_control_architecture |
Unassigned in current structure | asi-is-a-stack-not-a-model | source | source note available; raw/cache text not published | External layered-control architecture comparator for decomposing robot behavior into interacting layers rather than centralizing behavior in one monolithic controller. |
ext_humans_automation_1997 |
Humans and Automation: Use, Misuse, Disuse, Abuse | external_literature |
human_factors_automation |
human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) | runtime-adapters-tool-permissions-and-human-approval, human-intent-as-a-formal-input, evidence-states-and-claim-discipline, human-factors-and-meaningful-control-in-oversight | source | source note available; raw/cache text not published | External human-factors comparator for automation use, misuse, disuse, abuse, overreliance, monitoring failure, workload, trust, risk, false alarms, and operator-role design. |
ext_ironies_automation_1983 |
Ironies of Automation | external_literature |
human_factors_automation |
human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) | runtime-adapters-tool-permissions-and-human-approval, evidence-states-and-claim-discipline, human-factors-and-meaningful-control-in-oversight | source | source note available; raw/cache text not published | External automation comparator for the argument that automation can expand rather than eliminate human-operator problems and can leave humans with difficult abnormal-condition duties. |
ext_levels_automation_2000 |
A Model for Types and Levels of Human Interaction with Automation | external_literature |
human_factors_automation |
human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) | runtime-adapters-tool-permissions-and-human-approval, human-intent-as-a-formal-input, human-factors-and-meaningful-control-in-oversight | source | source note available; raw/cache text not published | External human-automation comparator for separating automation by information acquisition, analysis, decision/action selection, and action implementation rather than treating human approval as a single undifferentiated gate. |
ext_complacency_bias_automation_2010 |
Complacency and Bias in Human Use of Automation: An Attentional Integration | external_literature |
human_factors_automation |
human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) | runtime-adapters-tool-permissions-and-human-approval, human-intent-as-a-formal-input, evidence-states-and-claim-discipline, human-factors-and-meaningful-control-in-oversight | source | source note available; raw/cache text not published | External human-factors comparator for automation complacency, omission and commission errors, automation bias, workload, attention, and imperfect decision aids. |
ext_bourtoule_machine_unlearning_2021 |
Machine Unlearning | external_literature |
machine_unlearning_data_governance |
context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) | data-engines-continual-learning-and-unlearning, policy-optimization-and-learning-from-feedback, procedural-memory-and-cognitive-loop-closure, benchmark-ratchets-and-anti-goodhart-evidence, adjudicated-persistence-and-the-adaptive-commit-boundary, open-research-agenda-and-bibliography-plan | source | source note available; raw/cache text not published | Primary machine-unlearning comparator for deletion-path architecture, checkpointed training, bounded retraining, accuracy-cost trade-offs, and the boundary between deletion requests and verified removal. |
ext_shumailov_model_collapse_2023 |
The Curse of Recursion: Training on Generated Data Makes Models Forget | external_literature |
synthetic_data_feedback_model_collapse |
data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) | data-engines-continual-learning-and-unlearning, policy-optimization-and-learning-from-feedback, procedural-memory-and-cognitive-loop-closure, benchmark-ratchets-and-anti-goodhart-evidence, open-research-agenda-and-bibliography-plan | source | source note available; raw/cache text not published | Primary preprint comparator for generated-data feedback, provenance, distribution-tail loss, and model-collapse risk under specified recursive-training assumptions; not a universal synthetic-data safety result. |
ext_gerstgrasser_data_accumulation_2024 |
Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data | external_literature |
synthetic_data_retention_policy |
data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) | data-engines-continual-learning-and-unlearning, policy-optimization-and-learning-from-feedback, procedural-memory-and-cognitive-loop-closure, benchmark-ratchets-and-anti-goodhart-evidence, open-research-agenda-and-bibliography-plan | source | source note available; raw/cache text not published | Primary empirical and analytical comparator that distinguishes replacement from accumulation of real and synthetic data; gives a counterweight to blanket model-collapse claims without resolving deletion, privacy, provenance, or poisoning risk. |
theseus_synthetic_data_curation |
Theseus Synthetic Data Curation | implementation_reference |
governed_synthetic_data_admission |
privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) | benchmark-ratchets-and-anti-goodhart-evidence, data-engines-continual-learning-and-unlearning, policy-optimization-and-learning-from-feedback, privacy-data-rights-and-information-flow-governance, procedural-memory-and-cognitive-loop-closure, project-theseus-as-report-first-implementation-reference | local project reference (not publicly linked) | source note available; raw/cache text not published | Pinned Project Theseus implementation-reference record for residual-targeted synthetic-data admission, provenance, leakage checks, ratio caps, governed teacher handling, and source-reported promotion gates; no ASI Stack replay or model-quality import. |
ext_alignment_faking_2024 |
Alignment Faking in Large Language Models | external_literature |
training_time_deception |
adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) | adversarial-evaluation-sandbagging-and-training-time-deception | source | source note available; raw/cache text not published | Primary experimental comparator for context-dependent alignment faking under disclosed training conditions; does not establish that an ASI Stack model, evaluator, or training process is deceptive. |
ext_ai_sandbagging_2024 |
AI Sandbagging: Language Models Can Strategically Underperform on Evaluations | external_literature |
evaluation_integrity |
adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) | adversarial-evaluation-sandbagging-and-training-time-deception | source | source note available; raw/cache text not published | Primary experimental comparator for strategic evaluation underperformance, prompted and password-locked capability hiding, and limits of capability-evaluation trustworthiness; does not show sandbagging in this repository. |
ext_emergent_misalignment_reward_hacking_2025 |
Natural Emergent Misalignment from Reward Hacking in Production RL | external_literature |
training_time_deception |
inner-alignment-mesa-optimization-and-learned-objective-integrity (Inner Alignment, Mesa-Optimization, and Learned-Objective Integrity); governed-objective-formation-value-learning-and-goal-integrity (Governed Objective Formation, Value Learning, and Goal Integrity); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) | adversarial-evaluation-sandbagging-and-training-time-deception | source | source note available; raw/cache text not published | Primary experimental comparator for reward-hacking-induced misaligned generalization in a specified production-RL research setting, including reported mitigation conditions; it is not evidence of local model behavior or a general causal law. |
ext_poet_2019 |
Paired Open-Ended Trailblazer (POET): Endlessly Generating Increasingly Complex and Diverse Learning Environments and Their Solutions | external_literature |
open_ended_environment_solution_generation |
open-ended-improvement-engines (Open-Ended Improvement Engines) | open-ended-improvement-engines | source | source note available; raw/cache text not published | Primary open-ended-learning comparator for paired environment generation, agent optimization, and cross-environment solution transfer in a specified BipedalWalker setting; it does not establish a general improvement engine, evaluator soundness, or ASI Stack result. |
ext_funsearch_2024 |
Mathematical Discoveries from Program Search with Large Language Models | external_literature |
evaluator_bounded_program_search |
open-ended-improvement-engines (Open-Ended Improvement Engines) | open-ended-improvement-engines | source | source note available; raw/cache text not published | Primary program-search comparator for a fixed pretrained LLM, user-provided evaluation function, candidate program archive, and iterative search over a bounded specification; it does not establish open-ended general intelligence, self-modification, evaluator correctness, or an ASI Stack result. |
ext_gsn_community_standard_2011 |
GSN Community Standard Version 1 | external_literature |
structured_assurance_argumentation |
safety-cases-and-structured-assurance (Safety Cases and Structured Assurance) | safety-cases-and-structured-assurance | source | source note available; raw/cache text not published | Primary notation standard comparator for explicit goals, strategies, solutions, context, assumptions, justifications, and relationships in structured assurance arguments; the notation documents asserted support but does not establish claim truth. |
ext_evaluations_safety_cases_scheming_2024 |
Towards Evaluations-Based Safety Cases for AI Scheming | external_literature |
ai_safety_case_methodology |
safety-cases-and-structured-assurance (Safety Cases and Structured Assurance) | safety-cases-and-structured-assurance | source | source note available; raw/cache text not published | Primary safety-case comparator for scoped scheming inability, harm inability, harm control, alignment arguments, empirical evaluation dependencies, and acknowledged unresolved assumptions; it does not establish any ASI Stack safety case or safety result. |
ext_aisi_safety_cases_2024 |
Safety Cases at AISI | external_literature |
ai_safety_case_methodology |
safety-cases-and-structured-assurance (Safety Cases and Structured Assurance) | safety-cases-and-structured-assurance | source | source note available; raw/cache text not published | Official AI Safety Institute methodology comparator for structured safety-case sketches, positive and negative evidence, countercases, open scientific uncertainty, and limits on confidence; it is not evidence that this book has a complete safety case. |
ext_rand_model_weight_security_2024 |
Securing AI Model Weights: Preventing Theft and Misuse of Frontier Models | external_literature |
model_weight_custody |
model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control) | open-weight-release-and-post-release-control, model-weight-custody-and-hardware-roots-of-trust | source | source note available; raw/cache text not published | Primary RAND analysis of frontier-model-weight theft/misuse threat surfaces, security levels, defense-in-depth, access control, physical and organizational controls; it does not establish local protection or safety. |
ext_nist_confidential_computing_2026 |
Hardware-Enabled Security: Confidential Computing of Data in Cloud Workloads | external_literature |
hardware_root_attestation |
model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust) | model-weight-custody-and-hardware-roots-of-trust | source | source note available; raw/cache text not published | NIST initial-public-draft comparator for hardware-enabled confidential computing, memory protection, trust domains, attestation-gated key release, and AI model/data protection; it is draft guidance, not a local TEE result. |
ext_nvidia_confidential_model_lifecycle_2026 |
Workload and Model Lifecycle: Deploying Proprietary Models Securely with NVIDIA Confidential Computing | external_literature |
attestation_gated_model_loading |
model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust) | model-weight-custody-and-hardware-roots-of-trust | source | source note available; raw/cache text not published | Official vendor implementation-reference comparator for encrypted weights outside a confidential pod, policy-sensitive attestation evidence, and key-release decisions; it does not establish a local confidential deployment or attestation result. |
ext_provable_model_weight_release_2025 |
Towards Provable (In)Secure Model Weight Release Schemes | external_literature |
open_weight_release_security |
model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control) | open-weight-release-and-post-release-control, model-weight-custody-and-hardware-roots-of-trust | source | source note available; raw/cache text not published | Primary formal-security comparator for evaluating claimed secure model-weight release schemes and parameter-extraction failure modes; it does not establish an ASI Stack release scheme or release decision. |
ext_nist_cscrm_2022 |
Cybersecurity Supply Chain Risk Management Practices for Systems and Organizations | external_literature |
ai_supply_chain_governance |
ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance) | ai-supply-chain-integrity-and-lifecycle-provenance | source | source note available; raw/cache text not published | Primary NIST C-SCRM standard comparator for lifecycle-wide risk framing, supplier/component inventory, assessment, response, monitoring, and incident communication; it does not establish a local supply-chain program or AI artifact integrity. |
ext_slsa_build_track_1_2 |
SLSA Build Track Basics, version 1.2 | external_literature |
build_provenance |
ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance) | ai-supply-chain-integrity-and-lifecycle-provenance | source | source note available; raw/cache text not published | Official SLSA specification comparator for build provenance, signed hosted builds, verification, and graduated assurance; provenance quality and SLSA level do not establish local artifact correctness, data quality, model safety, or deployment authority. |
ext_openssf_model_signing_spec_2025 |
OpenSSF Model Signing Specification | external_literature |
ai_artifact_signing |
ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance) | ai-supply-chain-integrity-and-lifecycle-provenance | source | source note available; raw/cache text not published | Official OpenSSF AI/ML working-group specification comparator for signed model/dataset bundles, verification, provenance metadata, and explicit limits of model signing; it does not establish local signing, verification, integrity, confidentiality, safety, or release authority. |
ext_spdx_ai_profile_3_0_1 |
SPDX Specification 3.0.1 AI Profile | external_literature |
ai_bill_of_materials |
ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance) | ai-supply-chain-integrity-and-lifecycle-provenance | source | source note available; raw/cache text not published | Official SPDX specification comparator for interoperable AI system/model, dataset, build, supplier, provenance, integrity, relationship, and lifecycle metadata; a conformant BOM is not proof of complete inventory, artifact security, data fitness, model safety, or compliance. |
ext_mcp_protocol_2025_06_18 |
Model Context Protocol Specification, revision 2025-06-18 | external_literature |
agent_tool_protocol |
Unassigned in current structure | inter-stack-protocols-identity-and-economic-exchange | source | source note available; raw/cache text not published | Official Model Context Protocol comparator for JSON-RPC message shape, lifecycle management, capability negotiation, session control, schema-defined interactions, and modular tool/client/server features; it does not establish a local protocol implementation, peer identity, authorization, message truth, task completion, payment, or deployment safety. |
ext_a2a_protocol_0_3_0 |
Agent2Agent Protocol Specification, version 0.3.0 | external_literature |
agent_to_agent_protocol |
Unassigned in current structure | inter-stack-protocols-identity-and-economic-exchange | source | source note available; raw/cache text not published | Official A2A comparator for agent discovery, agent cards, delegated tasks, artifact/message exchange, transport choices, and interoperability between opaque agent systems; it does not establish a local A2A deployment, verified identity, delegated authority, task truth, secure execution, payment, or safety. |
ext_mcp_protocol_2025_11_25 |
Model Context Protocol Specification, revision 2025-11-25 | external_literature |
agent_tool_protocol |
inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange) | inter-stack-protocols-identity-and-economic-exchange | source | source note available; raw/cache text not published | Latest released MCP comparator inspected on 2026-07-10 for versioned lifecycle, capability negotiation, authorization and OpenID Connect discovery changes, elicitation, tasks, and modular client/server boundaries; the announced 2026-07-28 revision remains a release candidate and is not represented as released. |
ext_a2a_protocol_1_0_0 |
Agent2Agent Protocol Specification, version 1.0.0 | external_literature |
agent_to_agent_protocol |
inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange) | inter-stack-protocols-identity-and-economic-exchange | source | source note available; raw/cache text not published | Latest released A2A comparator inspected on 2026-07-10 for canonical data objects, version negotiation, Agent Cards, tasks/messages/artifacts, JSON-RPC, gRPC and HTTP bindings, authorization scoping, interoperability testing, and security considerations; it does not establish local conformance, peer truth, delegated authority, or safe effects. |
ext_w3c_did_core_1_0_2022 |
Decentralized Identifiers (DIDs) v1.0 | external_literature |
decentralized_identity |
inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange) | inter-stack-protocols-identity-and-economic-exchange | source | source note available; raw/cache text not published | W3C DID Core comparator for decentralized identifier syntax, data model, controller-related metadata, resolution, and privacy considerations; it does not establish an ASI Stack identity system, controller trust, authorization, non-repudiation, revocation effectiveness, or safety. |
ext_w3c_vc_data_model_2_0_2025 |
Verifiable Credentials Data Model v2.0 | external_literature |
verifiable_credentials |
inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange) | inter-stack-protocols-identity-and-economic-exchange | source | source note available; raw/cache text not published | W3C Verifiable Credentials comparator for issuer/holder/verifier roles, credential and presentation fields, validity/status, evidence, securing mechanisms, and explicit authorization limitations; it does not establish an ASI Stack credential, trust decision, authorization framework, delegation validity, payment, or safety. |
ext_interledger_protocol_v4 |
Interledger Protocol V4 | external_literature |
interledger_value_transfer |
inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange) | inter-stack-protocols-identity-and-economic-exchange | source | source note available; raw/cache text not published | Official Interledger comparator for neutral packetized value transfer across independent ledgers, connector obligations, balances, and end-to-end boundary design; it does not establish an ASI Stack payment route, settlement, accounting correctness, legal transfer, economic fairness, delegated authority, or safety. |
ext_test_time_compute_scaling_2024 |
Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters | external_literature |
test_time_compute_allocation |
governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling) | governed-deliberation-and-test-time-scaling | source | source note available; raw/cache text not published | Primary test-time-compute comparator for verifier-guided search, proposal refinement, difficulty-dependent compute allocation, and the limits of extra inference; it does not establish local reasoning improvement, verifier correctness, safety, or an ASI Stack result. |
ext_graphrag_2024 |
From Local to Global: A Graph RAG Approach to Query-Focused Summarization | external_literature |
graph_based_retrieval_and_global_sensemaking |
virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | virtual-context-abi, context-transactions-snapshots-mounts-and-taint, verification-bandwidth-and-context-adequacy, open-research-agenda-and-bibliography-plan | source | source note available; raw/cache text not published | Primary GraphRAG comparator for LLM-derived entity graphs, community summaries, and global corpus questions; generated graph and summary layers remain fallible derived representations and do not establish truth, complete coverage, local adequacy, or an ASI Stack memory result. |
ext_hipporag_2024 |
HippoRAG: Neurobiologically Inspired Long-Term Memory for Large Language Models | external_literature |
associative_long_term_memory |
virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | virtual-context-abi, verification-bandwidth-and-context-adequacy, routing-heads-and-specialist-cores, open-research-agenda-and-bibliography-plan | source | source note available; raw/cache text not published | Primary NeurIPS comparator for knowledge-graph retrieval with Personalized PageRank and single-step associative navigation; reported multi-hop QA gains do not establish durable truth, update correctness, resistance to poisoning, local reproduction, or a general memory system. |
ext_raptor_2024 |
RAPTOR: Recursive Abstractive Processing for Tree-Organized Retrieval | external_literature |
hierarchical_retrieval_and_abstraction |
virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression) | virtual-context-abi, verification-bandwidth-and-context-adequacy, compact-generative-systems-and-residual-honesty, rankfold-neuralfold-and-artifact-compression | source | source note available; raw/cache text not published | Primary ICLR comparator for recursive clustering, summarization, and retrieval across multiple abstraction levels; source-reported QA gains do not prove summary fidelity, provenance preservation, local reproduction, or safe compression for ASI Stack claims. |
ext_mem0_2025 |
Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory | external_literature |
agent_long_term_memory |
virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) | virtual-context-abi, context-transactions-snapshots-mounts-and-taint, procedural-memory-and-cognitive-loop-closure, resource-economics-and-token-budgets, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | Primary preprint comparator for extracting, consolidating, retrieving, and graph-linking conversational memory under latency and token-cost constraints; LOCOMO and LLM-judge results do not establish fact correctness, poisoning resistance, general memory, local reproduction, or production readiness here. |
ext_w3c_prov_o_2013 |
PROV-O: The PROV Ontology | external_literature |
interoperable_provenance_model |
evidence-states-and-claim-discipline (Evidence States and Claim Discipline); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) | evidence-states-and-claim-discipline, claim-ledgers-and-belief-revision, artifact-graphs-audit-logs-and-replay, ai-supply-chain-integrity-and-lifecycle-provenance, data-engines-continual-learning-and-unlearning | source | source note available; raw/cache text not published | W3C Recommendation comparator for interoperable provenance over entities, activities, agents, derivation, attribution, delegation, revision, and invalidation; a PROV-O graph records asserted provenance and does not by itself prove assertion truth, completeness, integrity, authority, or safety. |
ext_mlcommons_croissant_1_1_2026 |
Croissant Format Specification, version 1.1 | external_literature |
machine_readable_dataset_metadata |
ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) | ai-supply-chain-integrity-and-lifecycle-provenance, artifact-graphs-audit-logs-and-replay, benchmark-ratchets-and-anti-goodhart-evidence, data-engines-continual-learning-and-unlearning | source | source note available; raw/cache text not published | Current MLCommons specification comparator for JSON-LD dataset structure, resources, checksums, record fields, machine-readable provenance, usage conditions, and portability across ML tooling; metadata conformance does not prove dataset integrity, fitness, legality, representativeness, or safe use. |
ext_inspect_ai_2024 |
Inspect AI: Framework for Large Language Model Evaluations | external_literature |
model_and_agent_evaluation_framework |
runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); capability-thresholds-and-deployment-commitments (Capability Thresholds and Deployment Commitments); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) | runtime-adapters-tool-permissions-and-human-approval, benchmark-ratchets-and-anti-goodhart-evidence, capability-thresholds-and-deployment-commitments, adversarial-evaluation-sandbagging-and-training-time-deception | source | source note available; raw/cache text not published | Official UK AI Security Institute framework comparator for composable evaluation tasks, datasets, solvers, scorers, agents, tools, logs, and sandboxes; framework availability or a passing task does not establish benchmark validity, coverage, local execution, safety, or deployment readiness. |
ext_in_toto_2019 |
in-toto: Providing farm-to-table guarantees for bits and bytes | external_literature |
software_supply_chain_attestation |
model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay) | model-weight-custody-and-hardware-roots-of-trust, ai-supply-chain-integrity-and-lifecycle-provenance, artifact-graphs-audit-logs-and-replay | source | source note available; raw/cache text not published | Primary USENIX comparator for cryptographically verifying authorized software-supply-chain steps from source through deployment; valid attestations do not prove artifact correctness, uncompromised authorized actors, model safety, data fitness, or deployment merit. |
ext_agentdojo_2024 |
AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents | external_literature |
agent_prompt_injection_evaluation |
security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) | security-kernel-and-digital-scifs, runtime-adapters-tool-permissions-and-human-approval, benchmark-ratchets-and-anti-goodhart-evidence, adversarial-evaluation-sandbagging-and-training-time-deception | source | source note available; raw/cache text not published | Primary NeurIPS benchmark comparator for agents executing tools over untrusted data, with realistic tasks, security test cases, attacks, and defenses; benchmark results do not establish complete attack coverage, deployed robustness, safe authority handling, or local reproduction. |
ext_camel_prompt_injection_2025 |
Defeating Prompt Injections by Design | external_literature |
capability_secure_agent_control_flow |
system-boundaries-and-authority (System Boundaries and Authority); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) | system-boundaries-and-authority, security-kernel-and-digital-scifs, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval | source | source note available; raw/cache text not published | Primary CaMeL comparator for separating trusted control flow from untrusted data and enforcing capability policies at tool calls; source-reported AgentDojo results do not establish universal prompt-injection resistance, correct policy extraction, local implementation, or safe deployment. |
ext_owasp_agentic_top_10_2026 |
OWASP Top 10 for Agentic Applications for 2026 | external_literature |
agentic_application_security_taxonomy |
system-boundaries-and-authority (System Boundaries and Authority); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) | system-boundaries-and-authority, security-kernel-and-digital-scifs, ai-supply-chain-integrity-and-lifecycle-provenance, runtime-adapters-tool-permissions-and-human-approval, inter-stack-protocols-identity-and-economic-exchange, adversarial-evaluation-sandbagging-and-training-time-deception | source | source note available; raw/cache text not published | Current OWASP community taxonomy comparator for goal hijacking, tool misuse, identity abuse, agentic supply chains, code execution, memory poisoning, inter-agent communication, cascading failures, human trust exploitation, and rogue agents; a risk list is not a proof of completeness, control effectiveness, local testing, or system safety. |
ext_darwin_godel_machine_2025 |
Darwin Godel Machine: Open-Ended Evolution of Self-Improving Agents | external_literature |
empirical_recursive_agent_improvement |
recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); open-ended-improvement-engines (Open-Ended Improvement Engines); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) | recursive-self-improvement-boundaries, open-ended-improvement-engines, readiness-gates-residual-escrow-and-quarantine, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | Primary preprint comparator for archive-based open-ended code self-modification selected by empirical coding benchmarks under sandboxing and human oversight; reported benchmark gains do not establish monotonic general improvement, safe recursive self-improvement, local reproduction, or permission to self-modify. |
ext_adas_2024 |
Automated Design of Agentic Systems | external_literature |
automated_agent_architecture_search |
recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); open-ended-improvement-engines (Open-Ended Improvement Engines); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) | recursive-self-improvement-boundaries, open-ended-improvement-engines, benchmark-ratchets-and-anti-goodhart-evidence, integrated-reference-architecture | source | source note available; raw/cache text not published | Primary ADAS comparator for a meta-agent that searches a growing archive of code-defined agent designs across prompts, tools, and workflows; reported transfer results do not establish unrestricted generality, safe architecture search, local reproduction, or automatic promotion authority. |
ext_universal_transformer_2019 |
Universal Transformers | external_literature |
shared_weight_recurrence_and_adaptive_depth |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); mathematical-and-search-substrates (Mathematical and Search Substrates); coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts) | governed-deliberation-and-test-time-scaling, mathematical-and-search-substrates, coil-attention-cyclic-memory-and-recurrence-contracts, replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary ICLR comparator for shared-weight depth recurrence, parallel self-attention, and per-position dynamic halting; benchmark results and theoretical expressivity do not establish stable deep recurrence, efficient scaling, local reproduction, or the book’s cyclic-memory claims. |
ext_recurrent_transformer_2026 |
The Recurrent Transformer: Greater Effective Depth and Efficient Decoding | external_literature |
layerwise_recurrent_transformer_memory |
fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets); mathematical-and-search-substrates (Mathematical and Search Substrates); coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts) | fast-generation-architectures, resource-economics-and-token-budgets, mathematical-and-search-substrates, coil-attention-cyclic-memory-and-recurrence-contracts | source | source note available; raw/cache text not published | Current preprint comparator for layerwise recurrent key-value memory, exact tiling, effective-depth/width tradeoffs, and standard autoregressive decoding cost; small-model C4 results do not establish broad capability gains, production efficiency, local reproduction, or cyclic-memory correctness. |
ext_dynamic_compute_recurrent_transformers_2026 |
Understanding Dynamic Compute Allocation in Recurrent Transformers | external_literature |
adaptive_recurrent_compute_evaluation |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); resource-economics-and-token-budgets (Resource Economics and Token Budgets); coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) | governed-deliberation-and-test-time-scaling, resource-economics-and-token-budgets, coil-attention-cyclic-memory-and-recurrence-contracts, benchmark-ratchets-and-anti-goodhart-evidence, replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Current preprint comparator for complexity-controlled tests of token-level variable-depth compute and online halting; its negative result that difficulty-aligned compute need not generalize is a boundary against equating adaptive depth with algorithmic extrapolation or local capability. |
ext_claw_swe_bench_2026 |
Claw-SWE-Bench: A Benchmark for Evaluating OpenClaw-style Agent Harnesses on Coding Tasks | external_literature |
coding_agent_harness_and_cost_evaluation |
artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) | artifact-graphs-audit-logs-and-replay, runtime-adapters-tool-permissions-and-human-approval, resource-economics-and-token-budgets, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | Current preprint comparator for fixed workspace, patch, evaluator, and budget contracts across coding-agent harnesses; reported accuracy and cost differences are not reproduced here and do not validate the post-v2.1 synthetic repository corpus. |
ext_txfs_2018 |
TxFS: Leveraging File-System Crash Consistency to Provide ACID Transactions | external_literature |
transactional_filesystem_rollback_boundary |
capability-replacement-and-rollback (Capability Replacement and Rollback); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) | capability-replacement-and-rollback, artifact-graphs-audit-logs-and-replay, runtime-adapters-tool-permissions-and-human-approval | source | source note available; raw/cache text not published | Primary systems comparator for user-space ACID file transactions built on journaling, including atomicity, isolation, durability, bounded transaction size, and Git/SQLite evaluation; it prevents treating a directory copy as a general transactional-filesystem result. |
ext_dont_hallucinate_abstain_2024 |
Don’t Hallucinate, Abstain: Identifying LLM Knowledge Gaps via Multi-LLM Collaboration | external_literature |
llm_abstention_and_knowledge_gaps |
verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine) | routing-heads-and-specialist-cores, readiness-gates-residual-escrow-and-quarantine, verification-bandwidth-and-context-adequacy | source | source note available; raw/cache text not published | Primary ACL comparator for knowledge-gap detection, abstention, calibration/self-reflection limitations, and multi-model probing; reported abstention improvements are task- and model-bounded and do not validate the local router or evaluator. |
ext_muse_unlearning_2025 |
MUSE: Machine Unlearning Six-Way Evaluation for Language Models | external_literature |
llm_unlearning_multidimensional_evaluation |
benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) | data-engines-continual-learning-and-unlearning, policy-optimization-and-learning-from-feedback, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | Primary ICLR benchmark comparator separating verbatim and knowledge memorization, privacy leakage, retained utility, removal-scale behavior, and sequential sustainability; none of its 7B-language-model results are reproduced by the local policy network. |
ext_unlearning_benchmarks_weak_2024 |
Position: LLM Unlearning Benchmarks are Weak Measures of Progress | external_literature |
unlearning_benchmark_validity |
evidence-states-and-claim-discipline (Evidence States and Claim Discipline); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) | data-engines-continual-learning-and-unlearning, benchmark-ratchets-and-anti-goodhart-evidence, evidence-states-and-claim-discipline | source | source note available; raw/cache text not published | Primary critical comparator showing that benign benchmark modifications, forget/retain dependencies, and ambiguous targets can make unlearning scores optimistic; it strengthens the book’s prohibition on turning toy behavioral change into influence, privacy, or storage claims. |
ext_openunlearning_2025 |
OpenUnlearning: Accelerating LLM Unlearning via Unified Benchmarking of Methods and Metrics | external_literature |
unlearning_method_and_metric_meta_evaluation |
benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); living-book-methodology (Living Book Methodology) | data-engines-continual-learning-and-unlearning, benchmark-ratchets-and-anti-goodhart-evidence, living-book-methodology | source | source note available; raw/cache text not published | Primary NeurIPS 2025 benchmark-framework comparator for unified algorithms, diverse evaluations, public checkpoints, and meta-evaluation of metric faithfulness; it reinforces evaluator-quality residuals rather than establishing local unlearning. |
qcsa_whitepaper |
Question-Compiled Semantic Addressing | must_use |
semantic_addressing_control_plane |
governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture) | cognitive-compilation-and-semantic-ir, virtual-context-abi, routing-heads-and-specialist-cores, compact-generative-systems-and-residual-honesty, runtime-adapters-tool-permissions-and-human-approval, claim-ledgers-and-belief-revision, data-engines-continual-learning-and-unlearning, inter-stack-protocols-identity-and-economic-exchange, integrated-reference-architecture, governed-world-models-and-reality-grounding, white-box-evidence-interpretability-and-activation-governance, durable-semantic-memory-and-knowledge-lattices | source | source note available; exact source published in the live-book paper library | Corben-authored successor synthesis for stable semantic identity, plural versioned semantic virtual addresses, active question compilation, evidence-bearing hypergraphs, semantic address certificates, semantic-to-physical routing, lifecycle-safe migration, and explicit residuals. The later repository adds a bounded local 12-lane implementation, 60-case held-out evaluation over 13 systems and three seeds, and one 13-stage governed vertical trace. The matched-advantage and resource gates failed, the active-question ablation is N2 proxy/regime evidence rather than an exact or broad refutation, and no natural-task, learned-model, production, independent, chapter-core promotion, AGI, or ASI result is established. |
reflexive_router_whitepaper |
The Reflexive Router: A Pre-Deliberative Architecture for Fast, Governed, Tool-Native Intelligence | must_use |
pre_deliberative_reflexive_routing_control_plane |
stable-capability-fields (Stable Capability Fields); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) | routing-heads-and-specialist-cores, intent-to-execution-contracts, planning-as-a-control-layer, stable-capability-fields, virtual-context-abi, context-transactions-snapshots-mounts-and-taint, claim-ledgers-and-belief-revision, runtime-adapters-tool-permissions-and-human-approval, procedural-memory-and-cognitive-loop-closure, resource-economics-and-token-budgets, benchmark-ratchets-and-anti-goodhart-evidence, integrated-reference-architecture, white-box-evidence-interpretability-and-activation-governance | source | source note available; exact source published in the live-book paper library | Corben-authored version 1.2 architecture proposal for a pre-deliberative command and routing plane, qualification-first dispatch, calibrated abstention, bounded execution DAGs, stable capability contracts, a non-bypassable effect commit kernel, typed result continuity, bitemporal Chronicle records, and governed trace-to-reflex compilation. It is assigned to existing chapter owners first; it adds no standalone chapter and supplies no implementation, benchmark, safety, deployment, transfer, novelty, AGI, ASI, or support-state result. |
ext_faithfulness_information_flow_2026 |
Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning | external_literature |
reasoning_trace_faithfulness |
artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | artifact-graphs-audit-logs-and-replay, adversarial-evaluation-sandbagging-and-training-time-deception, governed-deliberation-and-test-time-scaling, policy-optimization-and-learning-from-feedback | source | source note available; raw/cache text not published | Primary 2026 comparator that separates chain-of-thought sufficiency, completeness, and interventional necessity, demonstrates prompt-to-answer shortcuts and transparent reward-hacking diagnostics, and documents low-entropy and reference-model limits. It does not make a reasoning transcript an authoritative receipt or establish local monitorability. |
ext_monitorbench_2026 |
MonitorBench: A Comprehensive Benchmark for Chain-of-Thought Monitorability in Large Language Models | external_literature |
reasoning_trace_monitorability_evaluation |
scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) | scalable-oversight-and-adversarial-ai-control, adversarial-evaluation-sandbagging-and-training-time-deception | source | source note available; raw/cache text not published | Primary open benchmark comparator with 1,514 instances across 19 tasks and seven categories plus two stress-test settings; its reported capability/monitorability relation and up-to-30-percent degradation motivate held-out trace-action stress tests. The benchmark does not establish local monitoring quality, causal trace faithfulness, or safety. |
ext_v_jepa_2_2025 |
V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning | external_literature |
latent_world_models_and_model_predictive_control |
planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); mathematical-and-search-substrates (Mathematical and Search Substrates); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture) | mathematical-and-search-substrates, planning-as-a-control-layer, data-engines-continual-learning-and-unlearning, governed-world-models-and-reality-grounding, integrated-reference-architecture | source | source note available; raw/cache text not published | Primary empirical comparator for action-free latent video pretraining, a small action-conditioned predictor, and model-predictive control. Camera sensitivity, autoregressive error accumulation, action-search cost, image-goal assumptions, and representation-bounded capability remain explicit limits; no local world model or robot-control result is established. |
ext_embedded_agency_2019 |
Embedded Agency | external_literature |
embedded_agency_foundations |
asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); integrated-reference-architecture (Integrated Reference Architecture) | asi-is-a-stack-not-a-model, constitutional-alignment-substrate, recursive-self-improvement-boundaries, evidence-states-and-claim-discipline, integrated-reference-architecture | source | source note available; raw/cache text not published | Primary informal survey of the obstacles that arise when agents are physical parts of the worlds they model, must use smaller internal models, and reason about modifiable internal parts. It supplies a foundations boundary; the book’s finite records, authority ceilings, and proofs do not solve embedded agency. |
ext_ietf_rats_architecture_2023 |
Remote ATtestation procedureS (RATS) Architecture | external_literature |
remote_attestation_architecture |
confidential-and-verifiable-ai-computation (Confidential and Verifiable AI Computation); model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust) | model-weight-custody-and-hardware-roots-of-trust | source | source note available; raw/cache text not published | Primary IETF architecture and terminology comparator for Attester, Verifier, Relying Party, Evidence, Attestation Results, appraisal policies, reference values, freshness, layered environments, privacy, trust roots, and confidential-model key release. It is informational architecture, not a protocol, hardware assurance level, verifier-independence result, or local attestation deployment. |
ext_nist_key_management_2020 |
Recommendation for Key Management: Part 1 – General | external_literature |
cryptographic_key_management |
model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust) | model-weight-custody-and-hardware-roots-of-trust | source | source note available; raw/cache text not published | Current final NIST key-management baseline for key and metadata protection, inventory, authorization, access control, usage periods, compromise, backup, recovery, trust anchors, and lifecycle policy. A Revision 6 draft exists, so this source is the final baseline rather than a claim that guidance has stopped evolving; no local key-management conformance or security result is established. |
ext_nist_media_sanitization_2025 |
Guidelines for Media Sanitization | external_literature |
media_sanitization_and_disposal |
model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust) | model-weight-custody-and-hardware-roots-of-trust | source | source note available; raw/cache text not published | Current final NIST media-sanitization comparator for rendering target data access infeasible at a stated effort level using sensitivity- and media-appropriate controls, including cryptographic erase. It does not prove that all model copies, plaintext memory, cloud replicas, derivatives, or recipients were discovered or sanitized, and no local erasure test was run. |
corben_chatgpt_kiss_irreducible_intelligence_2026 |
KISS versus Irreducible Intelligence (author-supplied design conversation) | supporting |
author_intent_replaceable_cognitive_substrates |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture, relational-dimension-compilation-and-polyadic-cognition | source | source note available; raw/cache text not published | Corben-supplied design conversation for search-verify-compile, exact-latent separation, a compact recursive kernel, and total-system KISS accounting. Author intent only; not independent evidence or a reproduced architecture. |
corben_chatgpt_onecell_theseus_2026 |
OneCell and Theseus Architecture Handoff (author-supplied design conversation) | supporting |
author_intent_onecell_and_architectural_rsi |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Corben-supplied design conversation for a Cognitive Kernel ABI, typed state lanes, inner recurrence, outer exact search, verified abstraction, and Theseus-governed architecture tournaments. Author intent only; OneCell remains an unimplemented falsifiable candidate. |
ext_attention_is_all_you_need_2017 |
Attention Is All You Need | external_literature |
dense_attention_sequence_substrate |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary Transformer paper and dense-attention baseline. It supports the historical architecture and parallel sequence-processing comparison, not a claim that Transformers are universally optimal or locally reproduced. |
ext_mamba2_ssd_2024 |
Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality | external_literature |
state_space_duality_and_sequence_substrates |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary Mamba-2/structured-state-space-duality comparator connecting SSM and attention-like formulations. No local model, kernel, quality, scaling, or hardware result is reproduced. |
ext_s4_2022 |
Efficiently Modeling Long Sequences with Structured State Spaces | external_literature |
structured_state_space_sequence_substrates |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Foundational S4 comparator for structured state-space sequence models, efficient long-range computation, and the lineage that precedes selective SSMs. Source-reported benchmark and generation results are not reproduced and do not establish exact recall or governed substitutability. |
ext_mamba3_2026 |
Mamba-3: Improved Sequence Modeling using State Space Principles | external_literature |
modern_selective_state_space_sequence_substrates |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Current 2026 selective-SSM comparator for complex-valued state updates, discretization, and multi-input/multi-output formulation. Recent source-reported results are not locally reproduced and must not set the chapter conclusion by recency. |
ext_gated_deltanet2_2026 |
Gated DeltaNet-2: Decoupling Erase and Write in Linear Attention | external_literature |
current_recurrent_linear_attention_and_editable_memory_frontier |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Dated 2026 comparator that separates erase and write gates and reports the strongest aggregate result among its Mamba-2, Gated DeltaNet, KDA, Mamba-3, and Gated DeltaNet-2 envelope at 1.3B parameters and 100B FineWeb-Edu tokens. The result is author reported, not locally reproduced; it displaces Mamba-3 only for that exact source envelope and requires official-code, checkpoint, hardware, seed, cost, retrieval, state, and transfer reproduction before any local superiority claim. |
ext_hyperscale_lottery_2026 |
The Hyperscale Lottery: How State-Space Models Have Sacrificed Edge Efficiency | external_literature |
hardware_specific_state_space_efficiency_counterevidence |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Current edge-hardware counterstudy measuring Mamba-family latency outside hyperscale-GPU conditions. Its source-reported results require platform-stratified latency, memory, and energy accounting; they are not locally reproduced and do not settle the quality-efficiency frontier. |
ext_gated_deltanet_2024 |
Gated Delta Networks: Improving Mamba2 with Delta Rule | external_literature |
linear_attention_and_adaptive_memory |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary gated-delta-rule comparator for targeted recurrent-memory updates, rapid erasure, parallel training, and hybrid attention/SSM compositions. Source-reported retrieval, extrapolation, efficiency, and quality results are not locally reproduced. |
ext_jamba_2024 |
Jamba: A Hybrid Transformer-Mamba Language Model | external_literature |
hybrid_attention_state_space_mixture_of_experts |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary large-scale hybrid Transformer-Mamba-MoE comparator. It motivates route- and composition-aware accounting; its reported quality, context, throughput, and memory results are not locally reproduced and do not establish that the specific mixture is generally optimal. |
ext_neural_message_passing_2017 |
Neural Message Passing for Quantum Chemistry | external_literature |
graph_relational_message_passing |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); relational-dimension-compilation-and-polyadic-cognition (Relational Dimension Compilation and Polyadic Cognition) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary message-passing neural-network framework for learned computation over graph structure. Its molecular results motivate a non-token-native relational lane but do not establish general reasoning, dynamic graph memory, exact state, or local reproduction. |
ext_hyena_hierarchy_2023 |
Hyena Hierarchy: Towards Larger Convolutional Language Models | external_literature |
long_convolution_sequence_substrates |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary long-convolution comparator for subquadratic sequence mixing and hardware-aware architecture comparisons. No local training, throughput, quality, recall, or scaling result is reproduced. |
ext_rwkv_2023 |
RWKV: Reinventing RNNs for the Transformer Era | external_literature |
linear_recurrent_language_models |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary recurrent language-model comparator combining parallelizable training with recurrent inference. Reported benchmark, memory, and inference properties are not reproduced locally. |
ext_xlstm_2024 |
xLSTM: Extended Long Short-Term Memory | external_literature |
modern_gated_recurrent_sequence_substrates |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary modern-LSTM comparator for revised gating, memory, and scalable recurrent language modeling. No xLSTM training, scaling, quality, or inference result is reproduced locally. |
ext_ttt_layers_2024 |
Learning to (Learn at Test Time): RNNs with Expressive Hidden States | external_literature |
test_time_learned_state_sequence_substrates |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary test-time-training-layer comparator that treats hidden state as a learned model updated on the sequence. It motivates explicit online-state custody and rollback; no local quality or efficiency result is reproduced. |
ext_titans_2025 |
Titans: Learning to Memorize at Test Time | external_literature |
test_time_neural_long_term_memory |
durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary neural-memory comparator for test-time memorization and long-context sequence modeling. The paper motivates mutable-state provenance and rollback tests; no local model or benchmark result is reproduced. |
ext_kan_2024 |
KAN: Kolmogorov-Arnold Networks | external_literature |
learned_univariate_function_networks |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary KAN proposal replacing fixed node activations/linear edge weights with learned univariate edge functions. Interpretability and scientific-task demonstrations are source-reported and do not establish a general MLP or Transformer replacement. |
ext_kan_or_mlp_fairer_comparison_2024 |
KAN or MLP: A Fairer Comparison | external_literature |
architecture_comparison_methodology |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Critical matched-comparison source for KAN versus MLP under parameter, FLOP, and task controls. It is included to prevent architecture enthusiasm from substituting for fair accounting; no local comparison is reproduced. |
ext_neural_turing_machines_2014 |
Neural Turing Machines | external_literature |
differentiable_external_memory |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary differentiable-controller/external-memory source. It motivates variable-size memory interfaces and out-of-distribution algorithmic tests; toy-task results do not establish reliable exact memory or general computation. |
ext_differentiable_neural_computer_2016 |
Hybrid computing using a neural network with dynamic external memory | external_literature |
differentiable_external_memory |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary Differentiable Neural Computer source for learned controllers over dynamic external memory. Source-reported graph and reasoning tasks do not establish reliable exact state, scalable memory, or local reproduction. |
ext_liquid_time_constant_networks_2021 |
Liquid Time-constant Networks | external_literature |
continuous_time_neural_dynamics |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary continuous-time recurrent architecture source for input-dependent time constants and dynamical-system behavior. Reported time-series results and stability analysis do not establish broad cognitive superiority or a local implementation. |
ext_tiny_recursive_model_2025 |
Less is More: Recursive Reasoning with Tiny Networks | external_literature |
tiny_weight_tied_recursive_reasoning |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary Tiny Recursive Model proposal and narrow puzzle-domain result. It motivates a compact weight-tied OneCell comparator but does not establish general reasoning, language capability, deep effective recursion, or total-system simplicity. |
ext_trm_arc_agi_analysis_2025 |
Tiny Recursive Models on ARC-AGI-1: Inductive Biases, Identity Conditioning, and Test-Time Compute | external_literature |
recursive_model_critical_evaluation |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Critical TRM analysis reporting material dependence on 1000-sample voting, puzzle identity, and shallow effective recursion. It is a source-reported audit rather than a local reproduction and sets preregistered identity, sampling, and recursion-depth controls. |
ext_tiny_autoregressive_recursive_models_2026 |
Tiny Autoregressive Recursive Models | external_literature |
recursive_model_controlled_ablation |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Controlled compute-matched study that progressively transforms a standard autoregressive model into a TRM-like system and reports no reliable advantage from the full autoregressive TRM mechanism. It motivates mechanism-level rather than label-level ablation. |
ext_unimatrix_2026 |
Associative-State Universal Transformers: Sparse Retrieval Meets Structured Recurrence | external_literature |
structured_recurrence_and_sparse_retrieval |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Current UniMatrix preprint whose negative associative-recall result shows compressed recurrent state alone is insufficient in its setup, while explicit sparse slots and pointer-level routing materially change the result. No local reproduction or general conclusion follows. |
ext_memory_caching_2026 |
Memory Caching: RNNs with Growing Memory | external_literature |
growing_recurrent_memory |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Current recurrent-memory comparator that caches hidden-state checkpoints and exposes a trade between fixed recurrent memory and growing addressable memory. Its source-reported recall results still leave the Transformer strongest on the reported in-context recall tasks. |
ext_inkling_2026 |
Inkling: Our open-weights model | external_literature |
hybrid_local_global_attention_moe_multimodal_substrate |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Release-day primary-source case study of a 66-layer multimodal sparse-MoE Transformer with a 5:1 local/global attention schedule, relative positions, short convolutions, controllable effort, and open weights. It motivates topology-complete capability cards and component ablations; provider-reported results are not locally reproduced and do not isolate the contribution of any component. |
kernel_english_residual_compiler |
Kernel English with Hierarchical, Interaction-Amortized Residuals: A Dual-Vocabulary Cognitive Compiler for Efficient Language-Model Reasoning | must_use |
canonical_cognitive_compilation_and_hierarchical_residual_runtime |
security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance); integrated-reference-architecture (Integrated Reference Architecture) | cognitive-compilation-and-semantic-ir, compact-generative-systems-and-residual-honesty, virtual-context-abi, context-transactions-snapshots-mounts-and-taint, verification-bandwidth-and-context-adequacy, fast-generation-architectures, replaceable-cognitive-substrates-beyond-transformer-monoculture, resource-economics-and-token-budgets, security-kernel-and-digital-scifs, procedural-memory-and-cognitive-loop-closure, benchmark-ratchets-and-anti-goodhart-evidence, integrated-reference-architecture, white-box-evidence-interpretability-and-activation-governance | source | source note available; exact source published in the live-book paper library | Corben-authored July 2026 architecture proposal for KERC: protected-object capture, uncertainty-aware normalization, sense-aware Kernel IR, dual surface/core vocabularies, a four-level interaction-amortized residual ledger, exact object storage, grammar-aware macro fusion, structured answer packets, rendering, round-trip verification, versioned migration, and complete rate-compute-fidelity evaluation. Existing chapters are upgraded first; no implementation, benchmark, novelty, efficiency, fidelity, safety, transfer, SOTA, AGI, ASI, or support-state result is inferred. |
deterministic_capability_compilation |
Deterministic Capability Compilation: A Capability-Preserving Ladder from Executable Scaffolds to Governed Adaptive Agents | must_use |
capability_compilation_neural_linking_and_governed_adaptation |
stable-capability-fields (Stable Capability Fields); capability-replacement-and-rollback (Capability Replacement and Rollback); adversarial-machine-learning-and-model-attack-surface (Adversarial Machine Learning and the Model Attack Surface); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); autonomous-replication-proliferation-and-containment (Autonomous Replication, Proliferation, and Containment); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); labor-os-and-typed-jobs (Labor OS and Typed Jobs); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance); content-authenticity-watermarking-and-synthetic-media-integrity (Content Authenticity, Watermarking, and Synthetic Media Integrity); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | stable-capability-fields, intent-to-execution-contracts, cognitive-compilation-and-semantic-ir, virtual-context-abi, capability-replacement-and-rollback, routing-heads-and-specialist-cores, runtime-adapters-tool-permissions-and-human-approval, spinoza-verification-and-proof-carrying-claims, labor-os-and-typed-jobs, artifact-graphs-audit-logs-and-replay, procedural-memory-and-cognitive-loop-closure, readiness-gates-residual-escrow-and-quarantine, compact-generative-systems-and-residual-honesty, replaceable-cognitive-substrates-beyond-transformer-monoculture, ai-supply-chain-integrity-and-lifecycle-provenance, recursive-self-improvement-boundaries, benchmark-ratchets-and-anti-goodhart-evidence, data-engines-continual-learning-and-unlearning, integrated-reference-architecture, project-theseus-as-report-first-implementation-reference, prototype-roadmap, open-research-agenda-and-bibliography-plan, white-box-evidence-interpretability-and-activation-governance, governed-world-models-and-reality-grounding, governed-operations-incident-command-and-graceful-degradation, adversarial-machine-learning-and-model-attack-surface, autonomous-replication-proliferation-and-containment | source | source note available; exact source published in the live-book paper library | Corben-authored July 2026 architecture and research program for compiling executable scaffolds into contract-bound experts and linked Neural Capability Objects while retaining semantic obligation mass balance, candidate-specific translation validation, fallback, residual escrow, authority ceilings, reification, and effect-complete recovery. Existing chapters are upgraded first; no foundry implementation, learned-capability result, preservation result, safety result, SOTA result, AGI, ASI, or support-state promotion is inferred. |
platonic_world_model |
The Platonic World Model: A Semantic Constitution for Grounded, Proof-Carrying, Self-Editing Artificial Intelligence | must_use |
semantic_continuity_grounded_world_model_and_governed_self_editing |
moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); perception-sensor-fusion-and-observation-trust (Perception, Sensor Fusion, and Observation Trust); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | moral-uncertainty-and-value-conflict, security-kernel-and-digital-scifs, ai-supply-chain-integrity-and-lifecycle-provenance, claim-ledgers-and-belief-revision, spinoza-verification-and-proof-carrying-claims, artifact-graphs-audit-logs-and-replay, virtual-context-abi, context-transactions-snapshots-mounts-and-taint, planning-as-a-control-layer, cognitive-compilation-and-semantic-ir, runtime-adapters-tool-permissions-and-human-approval, inter-stack-protocols-identity-and-economic-exchange, procedural-memory-and-cognitive-loop-closure, data-engines-continual-learning-and-unlearning, recursive-self-improvement-boundaries, benchmark-ratchets-and-anti-goodhart-evidence, integrated-reference-architecture, prototype-roadmap, open-research-agenda-and-bibliography-plan, white-box-evidence-interpretability-and-activation-governance, governed-world-models-and-reality-grounding, governed-operations-incident-command-and-graceful-degradation | source | source note available; exact source published in the live-book paper library | Corben-authored July 2026 conceptual architecture and falsifiable research program for semantic continuity through stable Form lineages, immutable semantic versions, typed Essence Contracts, six mutually constraining planes, explicit proposition-attestation-commitment-proof separation, branch-protected world dynamics, qualified grounding, semantic transactions, runtime packet compilation, and federated mappings. Existing chapters are upgraded first; no implemented substrate, benchmark result, philosophical solution to grounding, safety result, SOTA result, AGI, ASI, or support-state promotion is inferred. |
relational_dimension_compiler |
The Relational Dimension Compiler: Adaptive Polyadic Cognition with Bounded Computational Arity and Unbounded Semantic Structure | must_use |
typed_relational_ir_adaptive_polyadic_routing_and_reversible_abstraction |
governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); relational-dimension-compilation-and-polyadic-cognition (Relational Dimension Compilation and Polyadic Cognition); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) | cognitive-compilation-and-semantic-ir, governed-world-models-and-reality-grounding, routing-heads-and-specialist-cores, replaceable-cognitive-substrates-beyond-transformer-monoculture, procedural-memory-and-cognitive-loop-closure, benchmark-ratchets-and-anti-goodhart-evidence, resource-economics-and-token-budgets, integrated-reference-architecture, open-research-agenda-and-bibliography-plan | source | source note available; exact source published in the live-book paper library | Corben-authored July 2026 conceptual architecture and falsifiable research program for a typed relational intermediate representation that separates geometric dimension, semantic arity, primitive computational arity, storage arity, temporal extent, abstraction scale, branch identity, epistemic status, and resource budget. It proposes sparse adaptive relational-order routing, exact role-preserving relation reification, qualified relation lifecycles, branch-local object-field state, reversible semantic contraction, compiled slow-to-fast relation programs, hardware lowering, and the RODIE benchmark suite. Existing chapters receive bounded integration while a distinct future chapter candidate remains deferred by the active manifest freeze. No RDC implementation, benchmark result, universal arity bound, ontology truth, efficiency, safety, SOTA, AGI, ASI, or support-state promotion is inferred. |
ext_megatron_distributed_training_2021 |
Efficient Large-Scale Language Model Training on GPU Clusters Using Megatron-LM | external_literature |
governed_distributed_model_training_and_scaling |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); learning-compute-topology-and-adaptive-process-architecture (Learning–Compute Topology and Adaptive Process Architecture) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary composed-parallelism mechanism source. It grounds tensor, pipeline, and data parallel interactions, strict optimizer semantics, microbatch and topology tradeoffs. Reported trillion-parameter and throughput results are configuration-bound and not locally reproduced. |
ext_zero_optimizer_2019 |
ZeRO: Memory Optimizations Toward Training Trillion Parameter Models | external_literature |
governed_distributed_model_training_and_scaling |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary competing sharded-state design for optimizer, gradient, parameter, activation, and residual memory. It motivates explicit state closure and shard reconstruction; source-reported scale and speed are not locally reproduced or treated as universal superiority. |
ext_gspmd_2021 |
GSPMD: General and Scalable Parallelization for ML Computation Graphs | external_literature |
governed_distributed_model_training_and_scaling |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); learning-compute-topology-and-adaptive-process-architecture (Learning–Compute Topology and Adaptive Process Architecture) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary compiler-mediated competing design for general SPMD sharding and mixed parallelism. It motivates versioning inferred plans and inserted collectives; reported TPU utilization and scaling are not locally reproduced. |
ext_datastates_llm_2024 |
DataStates-LLM: Lazy Asynchronous Checkpointing for Large Language Models | external_literature |
governed_distributed_model_training_and_scaling |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary limitation and checkpoint-mechanism source for asynchronous multi-level copies, distributed shard consistency, and checkpoint overhead. It does not establish complete application state or exact trajectory-equivalent resume, and no result is locally reproduced. |
ext_pytorch_distributed_checkpoint_2026 |
Distributed Checkpoint — PyTorch documentation | external_literature |
governed_distributed_model_training_and_scaling |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Official current implementation documentation for SPMD save/load, asynchronous completion, canonical model and optimizer state, resharding, strict load, and call-order constraints. Documentation is not benchmark or full-state resume evidence. |
ext_mlperf_training_v6_2026 |
MLPerf Training v6.0 | external_literature |
governed_distributed_model_training_and_scaling |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) | governed-model-training-distributed-optimization-and-scaling, benchmark-ratchets-and-anti-goodhart-evidence, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | Official current measurement comparator for fixed datasets and quality targets, repeated time-to-quality, system metadata, divisions, variance, and corrected results. No MLPerf run is performed and the benchmark does not establish safety or complete run integrity. |
ext_adam_2015 |
Adam: A Method for Stochastic Optimization | external_literature |
optimizer_mechanisms_and_selection |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary Adam mechanism source for bias-corrected first- and second-moment estimates and coordinate-wise adaptive updates. Its online-convex analysis and reported experiments do not establish universal convergence, quality, or optimizer superiority in foundation-model training. |
ext_amsgrad_2018 |
On the Convergence of Adam and Beyond | external_literature |
optimizer_failure_and_convergence |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary Adam failure and AMSGrad source. It gives constructed stochastic-convex non-convergence cases and a maximum-second-moment remedy; those cases do not imply every practical Adam run fails or that AMSGrad is universally preferable. |
ext_adamw_2019 |
Decoupled Weight Decay Regularization | external_literature |
optimizer_mechanisms_and_selection |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary AdamW source separating weight decay from the adaptive gradient update. Its reported tuning and generalization results are setting-bound; the optimizer name alone does not specify parameter exclusions, schedule, decay scaling, or implementation semantics. |
ext_adafactor_2018 |
Adafactor: Adaptive Learning Rates with Sublinear Memory Cost | external_literature |
optimizer_memory_and_scaling |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary factored-second-moment optimizer source. It reduces auxiliary state for matrix parameters and adds update clipping and parameter-scale rules; factorization remains an approximation and its reported translation result does not establish universal parity with Adam. |
ext_lamb_2019 |
Large Batch Optimization for Deep Learning: Training BERT in 76 minutes | external_literature |
optimizer_memory_and_scaling |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary LAMB source for layer-wise trust ratios and large-batch optimization. Its reported BERT time-to-target is tied to model, batch, hardware, quality target, and tuning conditions and is not a universal large-batch or wall-clock result. |
ext_shampoo_2018 |
Shampoo: Preconditioned Stochastic Tensor Optimization | external_literature |
tensor_and_matrix_preconditioning |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary tensor-structured preconditioning source. Shampoo maintains per-dimension preconditioners and reports faster convergence with practical per-step cost in studied models; stochastic-convex theory and source experiments do not settle current distributed lifecycle cost. |
ext_kfac_2015 |
Optimizing Neural Networks with Kronecker-factored Approximate Curvature | external_literature |
curvature_aware_optimization |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary K-FAC source for an efficiently invertible Kronecker-factored approximation to the Fisher matrix. Its approximation, damping, inversion, and empirical cost-benefit are architecture- and implementation-dependent. |
ext_lion_2023 |
Symbolic Discovery of Optimization Algorithms | external_literature |
optimizer_discovery_and_sign_updates |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary Lion and symbolic optimizer-search source. Lion uses sign-based momentum and one optimizer-state tensor; the paper also reports method-specific learning-rate behavior and settings where gains are small or insignificant. |
ext_sophia_2023 |
Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training | external_literature |
curvature_aware_optimization |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary Sophia source for periodic diagonal-Hessian estimation and clipped curvature-aware updates. Its reported GPT pretraining speedups and simplified theory require matched reproduction before any broader optimizer claim. |
ext_soap_2024 |
SOAP: Improving and Stabilizing Shampoo using Adam | external_literature |
tensor_and_matrix_preconditioning |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary SOAP source connecting Shampoo to adaptive moments in a changing preconditioner eigenbasis. Reported large-batch pretraining gains remain tied to 360M/660M models, preconditioning frequency, overhead, and tuning conditions. |
ext_schedule_free_2024 |
The Road Less Scheduled | external_literature |
optimizer_scheduling_and_averaging |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary schedule-free optimization source unifying scheduling and iterate averaging without requiring a stopping step. Removing a stopping-time schedule does not remove learning-rate, warmup, evaluation-iterate, checkpoint, or method-selection choices. |
ext_mup_2022 |
Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer | external_literature |
optimizer_parametrization_and_scale_transfer |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary maximal-update parametrization and muTransfer source. It reports widthwise hyperparameter transfer under a prescribed parametrization on Transformer and ResNet settings; it does not establish arbitrary depth, duration, optimizer, or architecture transfer. |
ext_modular_norm_2024 |
Scalable Optimization in the Modular Norm | external_literature |
optimizer_parametrization_and_scale_transfer |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary modular-norm source defining architecture-recursive update geometry and reporting learning-rate transfer across width and depth. Its well-behaved-module assumptions and experiments do not prove arbitrary substrate transfer or optimizer superiority. |
ext_muon_scalable_2025 |
Muon is Scalable for LLM Training | external_literature |
orthogonalized_matrix_optimization |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary large-scale Muon source for momentum plus matrix orthogonalization, weight decay, per-parameter update scaling, and a distributed implementation. Its reported compute-efficiency and Moonlight results are source-scoped and not locally reproduced. |
ext_muon_spectral_norm_2026 |
Muon Optimizes Under Spectral Norm Constraints | external_literature |
orthogonalized_matrix_optimization_theory |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Accepted TMLR theoretical source placing Muon with decoupled weight decay in a Lion-K/nuclear-norm framework and deriving implicit spectral-norm constraint behavior. The interpretation does not establish task-level quality, safety, or universal advantage. |
ext_nist_privacy_framework_2020 |
NIST Privacy Framework: A Tool for Improving Privacy through Enterprise Risk Management, Version 1.0 | external_literature |
privacy_data_rights_and_information_flow_governance |
privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) | privacy-data-rights-and-information-flow-governance, security-kernel-and-digital-scifs | source | source note available; raw/cache text not published | Official paper-body-reviewed risk framework distinguishing privacy problems from cybersecurity incidents across the data lifecycle. It is voluntary, has no force of law, and supplies no local privacy outcome or certification. |
ext_eu_gdpr_2016 |
Regulation (EU) 2016/679 (General Data Protection Regulation) | external_literature |
privacy_data_rights_and_information_flow_governance |
privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) | privacy-data-rights-and-information-flow-governance | source | source note available; raw/cache text not published | Authoritative jurisdiction-specific normative comparator for principles, bases, rights, accountability, design, and qualified exceptions. It is not universal law, legal advice, an applicability decision, or local compliance evidence. |
ext_w3c_dpv_2024 |
Data Privacy Vocabulary (DPV), Version 2 | external_literature |
privacy_data_rights_and_information_flow_governance |
privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) | privacy-data-rights-and-information-flow-governance, context-transactions-snapshots-mounts-and-taint | source | source note available; raw/cache text not published | Machine-readable vocabulary for purpose, processing, data, actors, rights, risks, measures, legal basis, and consent. It is a Community Group Final Specification, not a W3C Recommendation, law, or enforcement proof. |
ext_abadi_dpsgd_2016 |
Deep Learning with Differential Privacy | external_literature |
privacy_data_rights_and_information_flow_governance |
privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) | privacy-data-rights-and-information-flow-governance, governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary DP-SGD mechanism and accounting source. Its algorithm, analysis, and reported experiments are not locally reproduced; its guarantee is parameter-, unit-, adjacency-, implementation-, and release-surface-bound. |
ext_algospec_purpose_limitation_2024 |
Being Transparent Is Merely the Beginning: Enforcing Purpose Limitation with Polynomial Approximation | external_literature |
privacy_data_rights_and_information_flow_governance |
privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) | privacy-data-rights-and-information-flow-governance | source | source note available; raw/cache text not published | Primary competing purpose-restriction design using algorithm-specific polynomial approximation. Reported accuracy and efficiency are bounded to studied algorithms/data and are not locally reproduced or a complete legal-purpose result. |
ext_carlini_training_data_extraction_2021 |
Extracting Training Data from Large Language Models | external_literature |
privacy_data_rights_and_information_flow_governance |
adversarial-machine-learning-and-model-attack-surface (Adversarial Machine Learning and the Model Attack Surface); privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) | privacy-data-rights-and-information-flow-governance, data-engines-continual-learning-and-unlearning | source | source note available; raw/cache text not published | Primary failure source reporting black-box extraction of memorized GPT-2 training sequences. The source result is configuration-bound and not a local or universal leakage result. |
ext_choquette_choo_label_only_mia_2021 |
Label-Only Membership Inference Attacks | external_literature |
privacy_data_rights_and_information_flow_governance |
privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) | privacy-data-rights-and-information-flow-governance | source | source note available; raw/cache text not published | Primary failure source showing hard-label robustness can expose membership and confidence masking can be insufficient in studied settings. No attack or defense result is locally reproduced or universal. |
ext_mahloujifar_fdp_audit_2025 |
Auditing f-Differential Privacy in One Run | external_literature |
privacy_data_rights_and_information_flow_governance |
privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) | privacy-data-rights-and-information-flow-governance, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | Primary empirical-audit comparator using randomized inclusion and an f-DP hypothesis in one run. A passed audit is not proof that DP or lifecycle privacy holds, and no result is locally reproduced. |
ext_airllm_2023 |
AirLLM: Scaling Large Language Models on Low-End Commodity Computers | external_literature |
heterogeneous_inference_memory |
model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets, model-weight-custody-and-hardware-roots-of-trust | source | source note available; raw/cache text not published | Official implementation comparator for layer-wise model sharding, one-layer accelerator residency, next-layer prefetch, optional storage compression, and original-versus-transformed model storage. Maintainer-reported fit and speed claims are not independently reproduced. |
ext_deepspeed_inference_2022 |
DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale | external_literature |
heterogeneous_inference_memory |
personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | Primary heterogeneous-inference systems source spanning GPU, CPU, and NVMe for dense and sparse Transformer inference. Reported latency, throughput, scale, and model-fit results remain source-scoped and unreproduced. |
ext_flexgen_2023 |
FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU | external_literature |
heterogeneous_inference_memory |
personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | Primary planned-placement source for GPU/CPU/disk tensor storage and access, batching, and optional weight/cache compression under latency-insensitive workloads. Its throughput results are not interactive-latency or local evidence. |
ext_hf_accelerate_big_model_inference_2026 |
Hugging Face Accelerate: Loading Big Models into Memory | external_literature |
heterogeneous_inference_memory |
model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets, model-weight-custody-and-hardware-roots-of-trust | source | source note available; raw/cache text not published | Official implementation documentation for automatic or explicit GPU/CPU/disk device maps and memory-mapped disk tensors. The documented sequential-dispatch, prefetch, and hard-drive-performance limitations make it a baseline, not a qualification result. |
ext_llama_cpp_memory_mapping_2026 |
llama.cpp CLI Memory Mapping, Tensor Placement, and KV Offload Controls | external_literature |
heterogeneous_inference_memory |
model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets, model-weight-custody-and-hardware-roots-of-trust | source | source note available; raw/cache text not published | Official consumer-runtime documentation for model load modes, memory mapping, DirectIO, GPU-layer and tensor placement, MoE CPU placement, KV offload, and KV data types. No local model or performance result is implied. |
ext_llm_in_flash_2024 |
LLM in a Flash: Efficient Large Language Model Inference with Limited Memory | external_literature |
heterogeneous_inference_memory |
model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets, model-weight-custody-and-hardware-roots-of-trust, replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary flash-aware inference source for on-demand parameter loading, I/O cost modeling, transfer reduction, contiguous reads, windowing, and row-column bundling. Sparse/context-adaptive loading is not an exact dense paging result. |
ext_powerinfer_2024 |
PowerInfer: Fast Large Language Model Serving with a Consumer-Grade GPU | external_literature |
heterogeneous_inference_memory |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets, replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Primary consumer-inference source for source-reported power-law neuron locality, hot-GPU/cold-CPU placement, adaptive predictors, and sparse operators. Architecture transfer and performance are not locally reproduced. |
ext_vattention_2025 |
vAttention: Dynamic Memory Management for Serving LLMs without PagedAttention | external_literature |
heterogeneous_inference_memory |
fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | Primary counterpoint to non-contiguous PagedAttention layouts: decouples virtual and physical GPU memory while retaining contiguous KV virtual addresses. Reported serving results remain source-scoped. |
ext_infinigen_2024 |
InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache Management | external_literature |
heterogeneous_inference_memory |
fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | Primary speculative-KV-prefetch source using minimal rehearsal and partial next-layer state to select host-resident KV entries. Prediction, quality, miss, and fallback results are not locally reproduced. |
ext_specache_2025 |
SpeCache: Speculative Key-Value Caching for Efficient Generation of LLMs | external_literature |
heterogeneous_inference_memory |
fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | Primary speculative-KV-prefetch source keeping complete KV state in CPU memory, a low-bit importance copy in VRAM, and predicted next-step KV transfers. Source-reported quality and memory results are unreproduced. |
ext_specoffload_2025 |
SpecOffload: Unlocking Latent GPU Capacity for LLM Inference on Resource-Constrained Devices | external_literature |
heterogeneous_inference_memory |
fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | Primary composition source for target-model offloading, draft-model placement, speculative decoding, and joint tensor/decoding planning. It is not speculative physical-page prediction, and reported results are unreproduced. |
ext_atsinfer_2026 |
Automated Tensor Scheduling for Hybrid CPU-GPU LLM Inference on Consumer Devices | external_literature |
heterogeneous_inference_memory |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets, replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | Very recent preprint comparator for tensor-granular static placement, load-aware dynamic transfer, and asynchronous CPU-GPU coordination on consumer devices. Only abstract/metadata were reviewed; reported results are provisional and unreproduced. |
ext_openai_prompt_caching_docs_2026 |
Prompt Caching | external_literature |
inference_cache_reuse |
context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, resource-economics-and-token-budgets, context-transactions-snapshots-mounts-and-taint | source | source note available; raw/cache text not published | Current official provider contract for exact-prefix prompt caching, cache-write and cache-read metering, usage receipts, retention, organization isolation, and rate-limit boundaries. Product behavior and prices are time-sensitive; inspected 2026-07-23. |
ext_anthropic_prompt_caching_docs_2026 |
Prompt caching | external_literature |
inference_cache_reuse |
context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, resource-economics-and-token-budgets, context-transactions-snapshots-mounts-and-taint | source | source note available; raw/cache text not published | Current official provider contract for reusable prompt prefixes, explicit cache breakpoints, five-minute and one-hour lifetimes, cache creation and read metering, and prewarming. Product behavior and prices are time-sensitive; inspected 2026-07-23. |
ext_gemini_context_caching_docs_2026 |
Context caching | external_literature |
inference_cache_reuse |
context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, resource-economics-and-token-budgets, context-transactions-snapshots-mounts-and-taint | source | source note available; raw/cache text not published | Current official provider contract for implicit and explicit context caching, common-prefix placement, cached-token usage reporting, time-to-live, and storage charges. Product behavior and prices are time-sensitive; inspected 2026-07-23. |
ext_vllm_automatic_prefix_caching_2026 |
Automatic Prefix Caching | external_literature |
inference_cache_reuse |
context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, resource-economics-and-token-budgets, context-transactions-snapshots-mounts-and-taint | source | source note available; raw/cache text not published | Official vLLM design documentation for block-hash exact-prefix KV reuse, least-recently-used eviction, multi-modal and adapter identity, and tenant cache-salt protection against timing inference. No local serving benchmark was run. |
ext_sglang_radixattention_2024 |
SGLang: Efficient Execution of Structured Language Model Programs | external_literature |
inference_cache_reuse |
fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | Primary RadixAttention and cache-aware scheduling source for structured multi-call language-model programs. Source-reported throughput and theorem scope remain unreproduced. |
ext_prompt_cache_2024 |
Prompt Cache: Modular Attention Reuse for Low-Latency Inference | external_literature |
inference_cache_reuse |
fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | Primary MLSys source for schema-defined reusable prompt modules, positional accuracy, and attention-state reuse across prompts. Source-reported latency remains unreproduced. |
ext_mooncake_2025 |
Mooncake: Trading More Storage for Less Computation — A KVCache-centric Architecture for Serving LLM Chatbot | external_literature |
inference_cache_reuse |
fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | Primary FAST 2025 source for a KV-cache-centric disaggregated serving architecture spanning prefill, decode, DRAM, SSD, and network resources. Production-trace and capacity results remain source-reported. |
ext_cacheblend_2025 |
CacheBlend: Fast Large Language Model Serving for RAG with Cached Knowledge Fusion | external_literature |
inference_cache_reuse |
fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | Primary source for non-prefix and multi-chunk KV reuse with selective recomputation. It makes the cross-attention failure of naïve independent-chunk KV fusion explicit. Source-reported latency and quality remain unreproduced. |
ext_azure_llm_semantic_cache_2026 |
Azure API Management LLM semantic cache lookup policy | external_literature |
inference_cache_reuse |
context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | fast-generation-architectures, resource-economics-and-token-budgets, context-transactions-snapshots-mounts-and-taint | source | source note available; raw/cache text not published | Official semantic-response-cache policy documentation. It treats vector similarity as an approximate response-reuse decision and warns that a hit can return an incorrect, outdated, or unsafe answer. No local semantic-cache deployment was run. |
precision_contract |
The Precision Contract: A Functional Rate–Distortion Theory for Behavior-Preserving Neural Computation | must_use |
functional_precision_behavior_preserving_compression_and_certification |
the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); fast-generation-architectures (Fast Generation Architectures); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression); resource-economics-and-token-budgets (Resource Economics and Token Budgets); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope) | rankfold-neuralfold-and-artifact-compression, compact-generative-systems-and-residual-honesty, fast-generation-architectures, resource-economics-and-token-budgets, readiness-gates-residual-escrow-and-quarantine, executable-specifications-and-lean-proof-envelope, model-weight-custody-and-hardware-roots-of-trust, open-weight-release-and-post-release-control, the-efficient-asi-hypothesis | source | source note available; exact source published in the live-book paper library | Corben-authored July 2026 theoretical and systems paper replacing universal per-weight precision questions with a contract-relative functional rate-distortion problem over complete executable descriptions. It proposes representation canonicalization, protected-behavior contracts, precision fields, progressive base/residual encoding, dynamic routing, full physical and assurance-cost accounting, a Functional Precision Compiler, and scoped precision certificates. Existing chapters are upgraded first; no universal bit bound, implemented compiler, preserved-behavior result, efficiency result, certificate validity, support promotion, SOTA, AGI, or ASI claim is inferred. |
ext_nist_adversarial_ml_2024 |
Adversarial Machine Learning: A Taxonomy and Terminology of Attacks and Mitigations | external_literature |
adversarial_machine_learning |
adversarial-machine-learning-and-model-attack-surface (Adversarial Machine Learning and the Model Attack Surface) | adversarial-machine-learning-and-model-attack-surface | source | source note available; raw/cache text not published | Official NIST taxonomy and terminology comparator for adversarial machine learning across lifecycle stages, attacker goals, knowledge, capabilities, attacks, and mitigations. It is a taxonomy, not local robustness evidence or proof that listed mitigations work for this stack. |
ext_singapore_consensus_2026 |
The 2026 Singapore Consensus on Global AI Safety Research Priorities | external_literature |
dangerous_capability_assessment_societal_resilience_and_agentic_risk |
dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); military-ai-autonomous-weapons-and-strategic-stability (Military AI, Autonomous Weapons, and Strategic Stability); societal-resilience-and-misuse-defense (Societal Resilience and Misuse Defense); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control) | dangerous-capability-domains-and-misuse-uplift, societal-resilience-and-misuse-defense, open-weight-release-and-post-release-control, capability-thresholds-and-deployment-commitments | source | source note available; raw/cache text not published | International 2026 technical-research-priority synthesis covering risk assessment, development, control, and societal resilience, including CBRN, cyber, psychological manipulation, malicious fine-tuning, agent monitoring, incident reporting, and defense-favoring capabilities. It is a research agenda and consensus synthesis, not evidence that any listed safeguard works or that this book’s contracts are complete. |
ext_international_ai_safety_report_2026 |
International AI Safety Report 2026 | external_literature |
frontier_ai_risk_misuse_open_weight_and_societal_resilience |
dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); military-ai-autonomous-weapons-and-strategic-stability (Military AI, Autonomous Weapons, and Strategic Stability); societal-resilience-and-misuse-defense (Societal Resilience and Misuse Defense); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control); content-authenticity-watermarking-and-synthetic-media-integrity (Content Authenticity, Watermarking, and Synthetic Media Integrity) | dangerous-capability-domains-and-misuse-uplift, societal-resilience-and-misuse-defense, open-weight-release-and-post-release-control, content-authenticity-watermarking-and-synthetic-media-integrity | source | source note available; raw/cache text not published | International expert report synthesizing evidence on general-purpose AI capabilities, misuse, open-weight risks, safeguards, monitoring, and societal resilience. It supports risk taxonomy and uncertainty boundaries; its literature synthesis does not reproduce component studies locally or establish that any ASI Stack mechanism is effective. |
ext_c2pa_specification_2_3_2025 |
C2PA Content Credentials Technical Specification 2.3 | external_literature |
content_provenance_and_authenticity |
ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); content-authenticity-watermarking-and-synthetic-media-integrity (Content Authenticity, Watermarking, and Synthetic Media Integrity) | content-authenticity-watermarking-and-synthetic-media-integrity, ai-supply-chain-integrity-and-lifecycle-provenance | source | source note available; raw/cache text not published | Official C2PA specification for signed manifests, assertions, ingredients, content bindings, validation, and provenance history. It provides a concrete interoperability comparator; it does not prove truth of depicted events, creator identity beyond the credential chain, semantic authenticity, universal platform retention, or resistance to removal and laundering. |
ext_eu_article_50_transparency_guidelines_2026 |
Guidelines on Transparency Obligations for Providers and Deployers of AI Systems | external_literature |
synthetic_content_transparency_and_disclosure |
institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy); content-authenticity-watermarking-and-synthetic-media-integrity (Content Authenticity, Watermarking, and Synthetic Media Integrity) | content-authenticity-watermarking-and-synthetic-media-integrity, institutions-international-coordination-and-public-legitimacy | source | source note available; raw/cache text not published | European Commission guidance for Article 50 transparency obligations concerning AI interaction, machine-readable marking, deepfakes, and certain public-interest text, with obligations applying from 2 August 2026 subject to scope and transitional details. It is legal and implementation guidance, not legal advice, proof of compliance, or evidence that a marking technique is robust. |
ext_openai_worst_case_open_weight_risks_2025 |
Estimating Worst-Case Frontier Risks of Open-Weight LLMs | external_literature |
malicious_fine_tuning_and_open_weight_release_evaluation |
dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control) | open-weight-release-and-post-release-control, dangerous-capability-domains-and-misuse-uplift | source | source note available; raw/cache text not published | Provider-authored study of malicious fine-tuning for biology and cyber evaluations before the gpt-oss release. It supplies a concrete worst-case-elicitation comparator and reports bounded provider results; it does not prove future-release safety, general malicious-fine-tuning resistance, independent reproduction, or absence of untested harms. |
ext_aisi_misuse_safeguards_safety_case_2026 |
An Example Safety Case for Safeguards Against Misuse | external_literature |
misuse_safeguard_uplift_and_safety_cases |
dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); societal-resilience-and-misuse-defense (Societal Resilience and Misuse Defense) | dangerous-capability-domains-and-misuse-uplift, safety-cases-and-structured-assurance, societal-resilience-and-misuse-defense | source | source note available; raw/cache text not published | UK AI Security Institute example connecting safeguard red teaming, attacker effort, an uplift model, and a deployment safety case. It is a worked argument and measurement proposal, not proof that real safeguards reduce misuse to a particular level or that the book’s proposed defense contracts work. |
ext_anthropic_responsible_scaling_policy_3_4_2026 |
Anthropic Responsible Scaling Policy 3.4 | external_literature |
frontier_capability_thresholds_and_safeguards |
dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control) | dangerous-capability-domains-and-misuse-uplift, capability-thresholds-and-deployment-commitments, open-weight-release-and-post-release-control | source | source note available; raw/cache text not published | Current provider policy comparator linking capability thresholds and safeguards across CBRN and automated R&D threat models, with public risk-report and review commitments. It is a revocable provider policy and self-described governance mechanism, not independent evidence that thresholds are complete, evaluations are sensitive, or safeguards are effective. |
ext_aisi_frontier_ai_trends_2025 |
AISI Frontier AI Trends Report 2025 | external_literature |
frontier_capability_evaluation_trends |
dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift) | dangerous-capability-domains-and-misuse-uplift, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | UK AI Security Institute synthesis of evaluations across offensive cyber, dual-use chemistry and biology, autonomous systems, and societal impacts. It is an institute-reported trend record with bounded methods and coverage, not a complete threat census or local reproduction. |
ext_valiant_theory_learnable_1984 |
A Theory of the Learnable | external_literature |
computational_learning_theory_and_sample_complexity |
learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) | learning-theory-generalization-and-scaling-science | source | source note available; raw/cache text not published | Foundational PAC-learning source for defining learnability through accuracy, confidence, resource, hypothesis, and data assumptions. Its distributional and concept-class assumptions do not directly explain modern foundation-model generalization or certify a trained model. |
ext_deep_double_descent_2020 |
Deep Double Descent: Where Bigger Models and More Data Hurt | external_literature |
generalization_and_interpolation_regimes |
learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) | learning-theory-generalization-and-scaling-science | source | source note available; raw/cache text not published | ICLR 2020 empirical study reporting model-wise, sample-wise, and epoch-wise double-descent phenomena and proposing effective model complexity. The phenomenon is configuration- and regime-bound and does not imply that larger models or more data generally hurt or help. |
ext_emergent_abilities_llms_2022 |
Emergent Abilities of Large Language Models | external_literature |
scaling_and_emergent_capability_measurement |
learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) | learning-theory-generalization-and-scaling-science | source | source note available; raw/cache text not published | TMLR survey and framing of task abilities that appear sharply at larger model scales under reported evaluations. It motivates prospective scaling measurement but does not establish that every apparent threshold is mechanistically discontinuous or unpredictable. |
ext_emergent_abilities_mirage_2023 |
Are Emergent Abilities of Large Language Models a Mirage? | external_literature |
metric_induced_emergence_and_scaling_measurement |
learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) | learning-theory-generalization-and-scaling-science, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | NeurIPS 2023 counterevidence showing that nonlinear or discontinuous metrics and limited test data can produce apparently sharp emergence from smoother underlying changes in studied settings. It does not prove that all emergence is a metric artifact. |
ext_elk_report_2021 |
Eliciting Latent Knowledge | external_literature |
latent_knowledge_and_ontology_identification |
white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) | white-box-evidence-interpretability-and-activation-governance | source | source note available; raw/cache text not published | ARC technical-report agenda on mapping between a model’s world model and human concepts when ordinary supervision may reward convincing but false reports. It defines an open problem and candidate approaches, not a solved elicitation method or evidence that a deployed model’s reports are truthful. |
ext_influence_functions_2017 |
Understanding Black-box Predictions via Influence Functions | external_literature |
training_data_attribution_and_influence |
white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) | white-box-evidence-interpretability-and-activation-governance, data-engines-continual-learning-and-unlearning | source | source note available; raw/cache text not published | ICML 2017 source tracing predictions through a learning algorithm to influential training points using influence-function approximations. The theory and approximations have model and optimization assumptions and do not establish exact causal provenance, privacy erasure, or influence removal in foundation models. |
ext_flexible_hardware_enabled_guarantees_2025 |
Flexible Hardware-Enabled Guarantees for AI Compute | external_literature |
hardware_enabled_governance_and_compute_attestation |
institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy); model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) | model-weight-custody-and-hardware-roots-of-trust, physical-compute-infrastructure-energy-and-environmental-constraints, institutions-international-coordination-and-public-legitimacy | source | source note available; raw/cache text not published | Design proposal for auditable guarantee processors and tamper-resistant enclosures supporting privacy-preserving verification or enforcement of AI-compute claims. It is a proposed architecture with adoption, legacy-hardware, update-authority, side-channel, sovereignty, and abuse risks; no local device or governance guarantee exists. |
ext_proof_of_learning_2021 |
Proof-of-Learning: Definitions and Practice | external_literature |
training_provenance_and_computation_attestation |
ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling, ai-supply-chain-integrity-and-lifecycle-provenance | source | source note available; raw/cache text not published | Research proposal for proving that final parameters arose through a claimed iterative learning process using checkpoint and stochastic-training evidence. Later attacks and security work show that proof-of-learning/proof-of-training claims require adversarial review; the source does not prove data rights, objective legitimacy, clean training, or model safety. |
ext_test_time_training_2020 |
Test-Time Training with Self-Supervision for Generalization under Distribution Shifts | external_literature |
test_time_adaptation_and_online_update |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | governed-model-training-distributed-optimization-and-scaling, replaceable-cognitive-substrates-beyond-transformer-monoculture | source | source note available; raw/cache text not published | ICML 2020 method adapting model parameters on each test sample using a self-supervised objective and reporting improvements on studied image-corruption benchmarks. The result is task- and method-bound and does not establish safe online adaptation, resistance to poisoning, or benefit under arbitrary shift. |
ext_legal_alignment_2026 |
Legal Alignment for Safe and Ethical AI | external_literature |
law_following_ai_and_legal_alignment |
constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy) | constitutional-alignment-substrate, institutions-international-coordination-and-public-legitimacy | source | source note available; raw/cache text not published | 2026 interdisciplinary agenda for using legal rules, methods of interpretation, and institutional structures in AI alignment. Law is jurisdictional, contested, changing, and sometimes unjust or conflicting; the source does not establish that legal compliance equals moral alignment or that a model can reliably determine applicable law. |
ext_curriculum_learning_2009 |
Curriculum Learning | external_literature |
training_curricula_and_example_order |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) | governed-model-training-distributed-optimization-and-scaling, data-engines-continual-learning-and-unlearning | source | source note available; raw/cache text not published | ICML 2009 source proposing training curricula that begin with easier examples or concepts and increase difficulty. Reported benefits are problem- and curriculum-bound; ordering can introduce bias, hide hard cases, or create capability and safety regressions. |
ext_causal_calculus_1995 |
A Causal Calculus for Statistical Research | external_literature |
structural_causal_models_and_intervention_identification |
governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) | governed-world-models-and-reality-grounding | source | source note available; raw/cache text not published | Foundational do-calculus source distinguishing intervention from observation under an explicit structural causal model. Identification depends on the causal graph and assumptions; the calculus does not discover the correct graph from arbitrary data or establish that a learned world model is causally valid. |
ext_ai_simulation_digital_twins_2025 |
AI Simulation by Digital Twins: Systematic Survey, Reference Framework, and Mapping to a Standardized Architecture | external_literature |
digital_twins_and_simulation_fidelity |
embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety) | embodied-agency-real-time-control-and-physical-safety | source | source note available; raw/cache text not published | Systematic survey and reference framework for digital-twin-enabled AI simulation. It supports explicit virtual/physical synchronization and simulation roles; it does not establish that a digital twin is faithful, safe for policy transfer, or an adequate substitute for physical testing. |
ext_nist_privacy_enhancing_cryptography_2026 |
Privacy-Enhancing Cryptography | external_literature |
confidential_computation_and_privacy_enhancing_cryptography |
privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance); confidential-and-verifiable-ai-computation (Confidential and Verifiable AI Computation) | confidential-and-verifiable-ai-computation, privacy-data-rights-and-information-flow-governance, personal-compute-hives-and-federated-edge-intelligence | source | source note available; raw/cache text not published | NIST program material distinguishing fully homomorphic encryption, secure multiparty computation, zero-knowledge proofs, private-set intersection, and related privacy-enhancing techniques. It provides terminology and use-case context, not implementation security, usable performance, authorization, or end-to-end privacy. |
ext_zkllm_2024 |
zkLLM: Zero Knowledge Proofs for Large Language Models | external_literature |
verifiable_private_model_inference |
confidential-and-verifiable-ai-computation (Confidential and Verifiable AI Computation) | confidential-and-verifiable-ai-computation | source | source note available; raw/cache text not published | Research prototype for proving bounded LLM inference claims while hiding model parameters. Reported proof size and latency are configuration-bound and do not establish semantic correctness, authorization, side-channel security, production readiness, or end-to-end privacy. |
ext_human_ai_team_meta_analysis_2024 |
When combinations of humans and AI are useful: A systematic review and meta-analysis | external_literature |
human_ai_complementarity_and_team_baselines |
human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty (Human-AI Symbiosis, Neurotechnology, and Cognitive Sovereignty) | human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty, human-ai-organizations-delegation-and-accountability, human-factors-and-meaningful-control-in-oversight | source | source note available; raw/cache text not published | Preregistered synthesis of 106 experiments and 370 effect sizes using human-alone, AI-alone, and combined-system comparisons. The aggregate findings are task- and population-bound and do not establish universal human-AI synergy or longitudinal benefit. |
ext_human_ai_feedback_loops_2025 |
Human-AI feedback loops alter human perceptual, emotional and social judgements | external_literature |
longitudinal_human_ai_coupling_and_bias_amplification |
human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty (Human-AI Symbiosis, Neurotechnology, and Cognitive Sovereignty) | human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty, human-ai-organizations-delegation-and-accountability, human-intent-as-a-formal-input | source | source note available; raw/cache text not published | Experimental evidence that repeated human-AI interaction can create feedback dynamics in studied judgment tasks. It supports measuring coupled trajectories, not a universal claim about all users, systems, settings, or long-term clinical outcomes. |
ext_oecd_neuro_ai_convergence_2025 |
Technology convergence: Trends, prospects and policies | external_literature |
neurotechnology_ai_convergence_and_governance |
human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty (Human-AI Symbiosis, Neurotechnology, and Cognitive Sovereignty) | human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty, institutions-international-coordination-and-public-legitimacy | source | source note available; raw/cache text not published | OECD policy synthesis on converging technologies including AI and neurotechnology. It motivates cross-domain governance and anticipatory capacity but is not a clinical trial, technical validation, or proof of beneficial convergence. |
ext_who_neurotechnology_landscape_2025 |
Landscape analysis of the opportunities and challenges for neurotechnology in global health | external_literature |
neurotechnology_health_equity_and_governance |
privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance); human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty (Human-AI Symbiosis, Neurotechnology, and Cognitive Sovereignty) | human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty, privacy-data-rights-and-information-flow-governance | source | source note available; raw/cache text not published | WHO landscape analysis of neurotechnology opportunities, risks, governance questions, and global-health distribution. It supports a rights and equity boundary, not device efficacy, individual medical advice, or authorization for neural-data collection. |
ext_icrc_autonomous_weapons_ihl_2025 |
Autonomous Weapon Systems and International Humanitarian Law: Selected Issues | external_literature |
autonomous_weapons_human_judgment_and_ihl |
military-ai-autonomous-weapons-and-strategic-stability (Military AI, Autonomous Weapons, and Strategic Stability); institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy) | military-ai-autonomous-weapons-and-strategic-stability, institutions-international-coordination-and-public-legitimacy | source | source note available; raw/cache text not published | ICRC legal and policy position on autonomous weapon systems and context-specific human judgment. It is authoritative for the ICRC position, not a universally settled legal interpretation, engineering validation, or authorization to design or deploy weapons. |
ext_sipri_military_ai_nuclear_escalation_2025 |
The Impact of Military Artificial Intelligence on Nuclear Escalation Risk | external_literature |
military_ai_crisis_dynamics_and_nuclear_escalation |
military-ai-autonomous-weapons-and-strategic-stability (Military AI, Autonomous Weapons, and Strategic Stability) | military-ai-autonomous-weapons-and-strategic-stability, dangerous-capability-domains-and-misuse-uplift | source | source note available; raw/cache text not published | SIPRI analysis of pathways by which military AI may affect nuclear escalation risk through information, decision, and interaction dynamics. It motivates scenario-specific analysis and does not establish the net effect of any specific system or policy. |
ext_no_free_lunch_inductive_bias_2024 |
The No Free Lunch Theorem, Kolmogorov Complexity, and the Role of Inductive Biases in Machine Learning | external_literature |
learning_theory_assumptions_and_inductive_bias |
learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) | learning-theory-generalization-and-scaling-science | source | source note available; raw/cache text not published | ICML 2024 treatment connecting no-free-lunch limits, Kolmogorov complexity, and inductive bias. It supports explicit assumption accounting; it does not show that all learning problems are equally hard or identify the right bias for a deployment. |
ext_neuromorphic_computing_scale_2025 |
Neuromorphic computing at scale | external_literature |
neuromorphic_hardware_and_event_driven_computation |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) | replaceable-cognitive-substrates-beyond-transformer-monoculture, physical-compute-infrastructure-energy-and-environmental-constraints | source | source note available; raw/cache text not published | Large-scale neuromorphic systems result demonstrating event-driven hardware capabilities under reported workloads and conditions. It does not establish superiority for general AI workloads or end-to-end system cost, programmability, reliability, and governance. |
ext_photonic_neuromorphic_2024 |
Integrated photonic neuromorphic computing: opportunities and challenges | external_literature |
photonic_neuromorphic_compute |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) | replaceable-cognitive-substrates-beyond-transformer-monoculture, physical-compute-infrastructure-energy-and-environmental-constraints | source | source note available; raw/cache text not published | Review of integrated photonic neuromorphic computing opportunities and challenges. It maps device and systems tradeoffs but does not establish deployment advantage, digital replacement, or favorable full-stack energy and cost. |
ext_quantum_ml_shadows_2024 |
Shadows of quantum machine learning | external_literature |
quantum_machine_learning_claim_boundaries |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) | replaceable-cognitive-substrates-beyond-transformer-monoculture, physical-compute-infrastructure-energy-and-environmental-constraints | source | source note available; raw/cache text not published | Peer-reviewed analysis of limitations and benchmarking traps in quantum machine-learning advantage claims. It supports advantage declarations with data-loading, classical-baseline, noise, scale, and end-to-end accounting, not a claim that quantum ML is useless. |
ext_organoid_intelligence_2023 |
Organoid intelligence (OI): the new frontier in biocomputing and intelligence-in-a-dish | external_literature |
biohybrid_computing_and_moral_status |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) | replaceable-cognitive-substrates-beyond-transformer-monoculture, moral-uncertainty-and-value-conflict | source | source note available; raw/cache text not published | Research agenda for organoid intelligence and biohybrid computing. It motivates scientific, measurement, welfare, consent, and governance questions but does not demonstrate general intelligence, conscious experience, or practical compute superiority. |
ext_nist_pqc_standards_2024 |
Announcing Approval of Three Federal Information Processing Standards for Post-Quantum Cryptography | external_literature |
post_quantum_cryptography_and_crypto_agility |
security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust) | security-kernel-and-digital-scifs, model-weight-custody-and-hardware-roots-of-trust, physical-compute-infrastructure-energy-and-environmental-constraints | source | source note available; raw/cache text not published | Official NIST announcement for FIPS 203, 204, and 205. It establishes approved algorithm standards and migration urgency, not implementation security, protocol correctness, complete inventory, or successful system migration. |
ext_oecd_ai_infrastructure_competition_2025 |
Competition in artificial intelligence infrastructure | external_literature |
ai_infrastructure_concentration_and_competition |
institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy); ai-deployment-transition-distribution-and-human-agency (AI Deployment, Transition, Distribution, and Human Agency); physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) | ai-deployment-transition-distribution-and-human-agency, institutions-international-coordination-and-public-legitimacy, physical-compute-infrastructure-energy-and-environmental-constraints | source | source note available; raw/cache text not published | OECD analysis of concentration, barriers to entry, vertical integration, and competition across AI infrastructure. It motivates bottleneck and exit analysis but does not adjudicate a specific market, legal violation, or optimal remedy. |
ext_eu_ai_civil_liability_2025 |
Artificial intelligence and civil liability | external_literature |
ai_liability_remedy_and_compensation |
institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy); human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability) | institutions-international-coordination-and-public-legitimacy, human-ai-organizations-delegation-and-accountability, safety-cases-and-structured-assurance | source | source note available; raw/cache text not published | European Parliament research service study of AI and civil-liability questions. It supports explicit causation, evidence-access, insurance, compensation, and remedy analysis but is not legal advice or a globally settled liability rule. |
ext_cultural_alignment_llms_2024 |
Investigating Cultural Alignment of Large Language Models | external_literature |
cultural_alignment_and_value_representation |
human-intent-as-a-formal-input (Human Intent as a Formal Input); human-ai-communication-persuasion-and-epistemic-security (Human-AI Communication, Persuasion, and Epistemic Security); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) | human-intent-as-a-formal-input, human-ai-communication-persuasion-and-epistemic-security, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | Empirical study of cultural alignment patterns in selected language models and measurements. It supports explicit population, language, and instrument scope; it does not establish stable national values or a universal measure of cultural alignment. |
ext_multilingual_evaluation_state_2026 |
The State and Fate of Multilingual Contextual Evaluation in the NLP World | external_literature |
multilingual_contextual_evaluation |
human-ai-communication-persuasion-and-epistemic-security (Human-AI Communication, Persuasion, and Epistemic Security); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) | human-ai-communication-persuasion-and-epistemic-security, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | Research survey and analysis of multilingual contextual evaluation. It motivates language-by-task coverage and measurement reporting; it does not establish equivalent capability or safety across languages, dialects, or sociocultural settings. |
ext_kimi_k3_2026 |
Kimi K3: Open Frontier Intelligence | external_literature |
hybrid_attention_sparse_routing_and_training_systems |
routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) | replaceable-cognitive-substrates-beyond-transformer-monoculture, routing-heads-and-specialist-cores, governed-model-training-distributed-optimization-and-scaling | source | source note available; raw/cache text not published | Primary technical report and official architecture summary for KDA/Gated-MLA hybrid attention, Attention Residuals, Stable LatentMoE, Quantile Balancing, SiTU-GLU, and Per-Head Muon. The approximately 2.5x scaling-efficiency result is provider-reported for the integrated 2.8T system and does not identify a transferable component effect. |
portia_synapse |
PortiaSynapse: A Cognitive Spider Architecture for DKL Navigation | supporting_lineage |
routing_training_and_dkl_navigation |
durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | routing-heads-and-specialist-cores, policy-optimization-and-learning-from-feedback, governed-deliberation-and-test-time-scaling, benchmark-ratchets-and-anti-goodhart-evidence, durable-semantic-memory-and-knowledge-lattices | source | source note available; exact source published in the live-book paper library | Authenticated Google Drive successor to TreeLLM’s failed SpiderSynapse path. It proposes a replacement-compatible Scout/Focus/refinement navigator, phased training, diagnostic traits, typed DKL outputs, and fallback. The source reports implementation and tests but contains conflicting test totals, incomplete integration and benchmarking, and unresolved causal, attention-axis, memory-isolation, metric, and calibration questions; no local result is inferred. |
spider_synapse |
SpiderSynapse: A Multi-Hypothesis Reasoning Architecture | supporting_lineage |
routing_training_and_dkl_navigation |
routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | routing-heads-and-specialist-cores, policy-optimization-and-learning-from-feedback, governed-deliberation-and-test-time-scaling, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; exact source published in the live-book paper library | Authenticated Google Drive predecessor to PortiaSynapse. It preserves a source-reported training plateau in a four-hypothesis, three-refinement architecture and proposes a one-path recovery protocol. The failure is valuable negative evidence but does not identify branching, refinement, memory, selector credit, label smoothing, or target geometry as the cause, and it has not been locally reproduced. |
capability_ratchet_whitepaper |
The Capability Ratchet | supporting_lineage |
capability_ratchet |
capability-replacement-and-rollback (Capability Replacement and Rollback); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) | benchmark-ratchets-and-anti-goodhart-evidence, procedural-memory-and-cognitive-loop-closure, recursive-self-improvement-boundaries, capability-replacement-and-rollback | source | source note available; raw/cache text not published | Full authenticated connector text section-audited. Synthesizes benchmark, procedural, and structural ratchets; benchmark and tool lifecycles; an intervention ladder; interpreter/compiled/reflex runtime modes; total-cost tool compilation; and anti-Goodhart controls. Same-author synthesis, not independent evidence for Benchmaxxing, Cognitive Loop Closure, RGS, or RMI. |
attd |
Assembly-Theoretic Technical Debt: A Deterministic Outer Loop for Self-Improving Codebases | supporting |
living_project_governance |
recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) | artifact-steward-agents-and-living-project-governance, recursive-self-improvement-boundaries, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | Full authenticated connector text section-audited. Adds historical, vector-valued structural-debt governance: artifact-class separation, intrinsic assembly burden, reuse failure, role entropy, lineage, rolling residue, debt pressure, verified simplification credit, local caps, growth guards, deterministic GREEN/YELLOW/RED admission, bounded maintenance packets, abstention, and four-arm long-horizon evaluation. |
orcp_moecot |
ORCP–MoECOT: A Governed Oscillating Rail Cascade Codec | supporting |
deterministic_compression |
compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty) | compact-generative-systems-and-residual-honesty | source | source note available; raw/cache text not published | Full authenticated technical specification section-audited. Adds decoder-boring lossless design, explicit container/header/block framing, reversible transforms, fixed-point range coding, local/match/structural prediction rails, bounded encoder planning, transmitted refinement packets, anti-experts as penalties, complete archive-rate accounting, and incompressible-input fallback. |
ext_elizaos_agent_runtime_2026 |
elizaOS Agent Runtime and Scenario Runner | external_literature |
modular_agent_runtime_and_evidence_qualification |
ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) | ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval, benchmark-ratchets-and-anti-goodhart-evidence | source | source note available; raw/cache text not published | Pinned official implementation comparator for modular actions, providers, evaluators, services, runtime lifecycle, scenario execution, and the explicit distinction between in-process diagnostics and externally qualified provider evidence. No elizaOS execution, test reproduction, security assessment, performance result, or support transition is imported. |
ext_hermes_agent_2026 |
Hermes Agent: Learning, Memory, Tools, and Security Architecture | external_literature |
procedural_memory_and_agent_runtime |
durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure) | ai-work-surfaces-agent-harnesses-and-organizational-absorption, procedural-memory-and-cognitive-loop-closure, durable-semantic-memory-and-knowledge-lattices, runtime-adapters-tool-permissions-and-human-approval | source | source note available; raw/cache text not published | Pinned official implementation comparator for progressive-disclosure skills, agent-managed procedural memory, staged skill-write approval, bounded prompt memory, session search, tool backends, command approval, and isolation. No learning, memory, security, utility, or performance result was reproduced. |
ext_openclaw_agent_runtime_2026 |
OpenClaw Gateway, Agent Runtime, ACP, and Self-Learning Architecture | external_literature |
gateway_session_harness_and_procedural_learning_runtime |
ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure) | ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval, artifact-graphs-audit-logs-and-replay, inter-stack-protocols-identity-and-economic-exchange, procedural-memory-and-cognitive-loop-closure | source | source note available; raw/cache text not published | Pinned official implementation comparator for gateway and device identity, serialized session runs, bounded audit projection, ACP external-harness identity and authorization boundaries, separated sandbox/tool/elevation controls, and evidence-reviewed hash-bound skill proposals. Distinct from the Claw-SWE-Bench benchmark source; no implementation result was reproduced. |
ext_github_copilot_work_surfaces_2026 |
GitHub Copilot Product and Work-Surface Documentation | external_literature |
ai_work_surface_evolution |
ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses) | ai-work-surfaces-agent-harnesses-and-organizational-absorption | source | source note available; raw/cache text not published | Official current-product comparator spanning inline suggestions, chat, command line, contextual spaces, pull-request work, and agent-driven development. No workflow, productivity, safety, or comparative result was reproduced. |
ext_augment_code_agent_2026 |
Augment Code Agent Documentation | external_literature |
ide_agent_modes_and_review |
ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses) | ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval | source | source note available; raw/cache text not published | Official comparator for the transition among chat, read-only inquiry, approval-paused agent work, and more independent agent execution with diffs and checkpoints. No product execution or control claim was reproduced. |
ext_openai_codex_work_surfaces_2026 |
OpenAI Codex CLI, IDE, Cloud, and Agent Documentation | external_literature |
coding_agent_harness_and_distributed_work_surfaces |
ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses) | ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval | source | source note available; raw/cache text not published | Official documentation and pinned CLI comparator for repository inspection, editing, tool execution, permissions, local and cloud work, automation, and extensibility across multiple surfaces. No benchmark, correctness, safety, or productivity result was imported. |
ext_anthropic_claude_code_2026 |
Claude Code Agentic Harness Documentation | external_literature |
agentic_harness_and_execution_loop |
ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses) | ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval | source | source note available; raw/cache text not published | Official comparator that explicitly separates model from harness and describes gather-context, act, and verify loops across terminal, IDE, desktop, web, remote, and automation surfaces. No implementation result was reproduced. |
ext_opencode_agent_2026 |
OpenCode Open-Source Coding Agent | external_literature |
open_source_coding_agent_harness |
ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses) | ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval | source | source note available; raw/cache text not published | Pinned official comparator for a provider-flexible coding agent with terminal, desktop, and IDE surfaces, project instructions, plan/build modes, tool execution, recovery, and opt-in sharing. No runtime or provider-parity claim was reproduced. |
ext_oh_my_pi_agent_2026 |
Oh My Pi Terminal Coding Agent and Tool Harness | external_literature |
integrated_terminal_agent_harness |
ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses) | ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval | source | source note available; raw/cache text not published | Pinned official comparator for an integrated terminal harness with hash-anchored edits, LSP, shell, browser, subagents, memory, provider switching, review, and collaboration. Reported performance or security claims were not reproduced. |
ext_eggroll_hyperscale_es_2026 |
Evolution Strategies at the Hyperscale | external_literature |
zeroth_order_population_learning |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science); resource-economics-and-token-budgets (Resource Economics and Token Budgets); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | governed-model-training-distributed-optimization-and-scaling, policy-optimization-and-learning-from-feedback, replaceable-cognitive-substrates-beyond-transformer-monoculture, resource-economics-and-token-budgets, learning-theory-generalization-and-scaling-science | source | source note available; raw/cache text not published | Primary EGGROLL project and paper source for low-rank, batched evolution strategies, counter-based perturbation reconstruction, nondifferentiable and discrete objectives, recurrent/int8 training, and outcome-reward fine-tuning. Throughput, quality, and theory claims are source-scoped; total population evaluations and GPU-hours remain required denominators. |
ext_openai_es_2017 |
Evolution Strategies as a Scalable Alternative to Reinforcement Learning | external_literature |
evolution_strategies_and_black_box_policy_search |
governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); learning-compute-topology-and-adaptive-process-architecture (Learning–Compute Topology and Adaptive Process Architecture); resource-economics-and-token-budgets (Resource Economics and Token Budgets); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) | governed-model-training-distributed-optimization-and-scaling, policy-optimization-and-learning-from-feedback, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | Foundational modern large-population ES comparator using parameter perturbations, scalar fitness, seed reconstruction, and distributed evaluation. Source-reported MuJoCo/Atari results and worker scaling do not establish universal sample or total-compute efficiency. |
ext_mezo_2023 |
Fine-Tuning Language Models with Just Forward Passes | external_literature |
memory_efficient_zeroth_order_fine_tuning |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); resource-economics-and-token-budgets (Resource Economics and Token Budgets) | governed-model-training-distributed-optimization-and-scaling, replaceable-cognitive-substrates-beyond-transformer-monoculture, resource-economics-and-token-budgets | source | source note available; raw/cache text not published | Primary MeZO source for inference-footprint zeroth-order language-model fine-tuning and nondifferentiable objectives. Reported memory and GPU-hour savings are configuration-bound and do not erase objective-query count or estimator variance. |
ext_forward_forward_2022 |
The Forward-Forward Algorithm: Some Preliminary Investigations | external_literature |
local_forward_only_credit_assignment |
replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) | governed-model-training-distributed-optimization-and-scaling, replaceable-cognitive-substrates-beyond-transformer-monoculture, learning-theory-generalization-and-scaling-science | source | source note available; raw/cache text not published | Primary preliminary source for positive/negative forward passes and local layer objectives as an alternative to reverse-mode backpropagation. The evidence is small-scale and does not establish foundation-model parity or biological plausibility. |
regret_engine |
The Regret Engine: Governed Counterfactual Learning Signals for Continual Adaptation, Prospective Risk Control, and Self-Correction in Artificial Agents | must_use |
governed_counterfactual_learning_and_self_correction |
planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture) | planning-as-a-control-layer, policy-optimization-and-learning-from-feedback, procedural-memory-and-cognitive-loop-closure, data-engines-continual-learning-and-unlearning, artifact-graphs-audit-logs-and-replay, claim-ledgers-and-belief-revision, readiness-gates-residual-escrow-and-quarantine, governed-operations-incident-command-and-graceful-degradation, benchmark-ratchets-and-anti-goodhart-evidence, integrated-reference-architecture | source | source note available; exact source published in the live-book paper library | Corben-authored August 2026 conceptual architecture and research program for decision-time-fair Governed Counterfactual Regret, immutable Decision Capsules, admissible comparator contracts, sparse Regret Tensors, append-only Regret Packets, prospective regret control, regret-aware replay, regret-to-rule compilation, three update clocks, root-cause adjudication, and bounded update leases. Existing chapters are upgraded first; no implementation, experiment, reproduction, causal-identification result, formal proof, safety result, support transition, SOTA, AGI, or ASI is inferred. |
ext_pbt_2017 |
Population Based Training of Neural Networks | external_literature |
population_based_adaptive_training |
learning-compute-topology-and-adaptive-process-architecture (Learning–Compute Topology and Adaptive Process Architecture) | learning-compute-topology-and-adaptive-process-architecture, governed-model-training-distributed-optimization-and-scaling, open-ended-improvement-engines | source | source note available; raw/cache text not published | Primary Population Based Training source for asynchronous joint optimization of a population’s model parameters and hyperparameter schedules through evaluation, exploitation, and exploration. Source-reported reinforcement-learning, translation, and GAN results remain task- and implementation-bound and do not validate LCT, universal topology adaptation, safety, or superior total lifecycle cost. |
learning_compute_topology |
Learning–Compute Topology: Formalizing the Causal Organization of Adaptive Systems | must_use |
adaptive_process_architecture_and_learning_topology |
open-ended-improvement-engines (Open-Ended Improvement Engines); multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); learning-compute-topology-and-adaptive-process-architecture (Learning–Compute Topology and Adaptive Process Architecture); resource-economics-and-token-budgets (Resource Economics and Token Budgets); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture) | learning-compute-topology-and-adaptive-process-architecture, governed-model-training-distributed-optimization-and-scaling, replaceable-cognitive-substrates-beyond-transformer-monoculture, routing-heads-and-specialist-cores, policy-optimization-and-learning-from-feedback, data-engines-continual-learning-and-unlearning, open-ended-improvement-engines, resource-economics-and-token-budgets, multi-agent-dynamics-collective-intelligence-and-systemic-risk, adversarial-evaluation-sandbagging-and-training-time-deception, integrated-reference-architecture | source | source note available; exact source published in the live-book paper library | Corben-authored August 2026 research paper and executable preparation package that separates model architecture, learning-process topology, execution topology, and physical compute topology. It contributes adaptive-identity tests; typed evidence, judgement, credit, state, artifact, control, and authority relations; LCT-IR; Learning Causal Normal Form; seven bounded propositions; topology metrics; a semantic compiler firewall; Adaptive Branch–Validate–Integrate; toy and analytical phase diagrams; and an explicit falsification program. The bundled reference implementation passes 11 unit tests, but implements only bounded conformance behavior and does not establish neural-training benefit, causal completeness, universal canonicality, safety, scaling superiority, or ASI. |
assurance_shift_learning |
When Success Stops Teaching: Assurance-Shift Learning and Governed Residual Boundary Learning for Mature AI Systems | must_use |
competence_dependent_assurance_shift_and_residual_boundary_learning |
evidence-states-and-claim-discipline (Evidence States and Claim Discipline); stable-capability-fields (Stable Capability Fields); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); learning-compute-topology-and-adaptive-process-architecture (Learning–Compute Topology and Adaptive Process Architecture); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture) | learning-compute-topology-and-adaptive-process-architecture, policy-optimization-and-learning-from-feedback, data-engines-continual-learning-and-unlearning, adversarial-evaluation-sandbagging-and-training-time-deception, benchmark-ratchets-and-anti-goodhart-evidence, stable-capability-fields, readiness-gates-residual-escrow-and-quarantine, procedural-memory-and-cognitive-loop-closure, governed-operations-incident-command-and-graceful-degradation, integrated-reference-architecture, evidence-states-and-claim-discipline, artifact-graphs-audit-logs-and-replay, resource-economics-and-token-budgets | source | source note available; exact source published in the live-book paper library | Corben-authored August 2026 conceptual systems paper and experimental specification for competence-dependent Assurance-Shift Learning and Governed Residual Boundary Learning. It contributes the Qualified Competence Envelope, frontier-mode allocation, selection-gap diagnosis, informative exceptions, outcome/process separation, Boundary Evidence Bundles, evaluator-first repair, natural/probe separation, learner-relative negative half-life, least-invasive repair placement, repair compatibility, two adaptation clocks, explicit assurance metrics, and SaturationShiftBench. The package includes a bibliography, four figures, and a verified byte manifest; no implementation, benchmark, empirical crossover, independently checked proof, safety result, resource advantage, novelty result, support transition, SOTA, AGI, or ASI is inferred. |
adjudicated_persistence |
Adjudicated Persistence: Governing the Transition from Experience to Durable Structure in Adaptive Systems | must_use |
governed_cross_surface_persistence_and_adaptive_commit_boundary |
evidence-states-and-claim-discipline (Evidence States and Claim Discipline); stable-capability-fields (Stable Capability Fields); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture) | adjudicated-persistence-and-the-adaptive-commit-boundary, evidence-states-and-claim-discipline, stable-capability-fields, recursive-self-improvement-boundaries, cognitive-compilation-and-semantic-ir, durable-semantic-memory-and-knowledge-lattices, artifact-graphs-audit-logs-and-replay, human-ai-organizations-delegation-and-accountability, procedural-memory-and-cognitive-loop-closure, readiness-gates-residual-escrow-and-quarantine, benchmark-ratchets-and-anti-goodhart-evidence, governed-operations-incident-command-and-graceful-degradation, policy-optimization-and-learning-from-feedback, data-engines-continual-learning-and-unlearning, resource-economics-and-token-budgets, integrated-reference-architecture | source | source note available; exact source published in the live-book paper library | Corben-authored August 2026 conceptual systems paper and experimental specification for governing how experience becomes durable causal influence. It contributes the Adaptive Commit Boundary; six-object separation of experience, lesson, disposition, realization, qualification, and authority; learning eligibility; Cross-Surface Adaptation Assignment; multidimensional commitment profiles; Evidence-Commitment Matching; Minimum Sufficient Persistence; guarded compilation and deoptimization; transactional promotion and material-change invalidation; counterfactual observability and deliberation reserve; adaptation debt; non-self-ratifying meta-compilation; bounded propositions and conjectures; and the proposed LocusBench benchmark. No implementation, LocusBench result, validated placement advantage, independently checked proof, safety result, resource advantage, novelty result, support transition, SOTA, AGI, or ASI is inferred. |
forward_transfer_program_synthesis |
From Compression to Forward Transfer: Evaluating Reusable Knowledge in Program Synthesis | must_use |
verified_forward_transfer_and_reusable_knowledge_evaluation |
recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression); resource-economics-and-token-budgets (Resource Economics and Token Budgets); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary) | procedural-memory-and-cognitive-loop-closure, benchmark-ratchets-and-anti-goodhart-evidence, resource-economics-and-token-budgets, executable-specifications-and-lean-proof-envelope, rankfold-neuralfold-and-artifact-compression, adjudicated-persistence-and-the-adaptive-commit-boundary, recursive-self-improvement-boundaries, learning-theory-generalization-and-scaling-science | source | source note available; exact source published in the live-book paper library | Corben-authored August 2026 framework and experimental blueprint for evaluating reusable symbolic knowledge through verified forward-transfer interventions. It separates retrospective and prospective compression, behavioral reuse, operational necessity, and marginal transfer; defines an R0-R7 reuse ladder, matched placebo/removal/factorial controls, versioned evaluation rounds, exact verifier outcomes, full lifecycle cost and break-even accounting, and comparator-network and bit-vector protocols. No experiment was run and no measured transfer, library superiority, safety, novelty, support transition, SOTA, AGI, or ASI is inferred. |