Skip to main content

Appendix A — Source Matrix

This matrix is generated from sources/source_inventory.json and dynamic chapter assignments in book_structure.json.

Source status is deliberately conservative. A source note means the source has been mined for drafting context; it does not by itself promote any chapter claim above argument.

ID Title Priority Layer Current dynamic assignments Original packet targets URL Current status Notes
ext_probe_control_tasks_2019 Designing and Interpreting Probes with Control Tasks external_literature mechanistic_interpretability white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) white-box-evidence-interpretability-and-activation-governance source source note available; raw/cache text not published Primary probe-method comparator for control tasks and selectivity: a probe must be evaluated against its capacity to learn control labels rather than treating linguistic-task accuracy as representation evidence. The source studies ELMo linguistic probes; it does not establish a universal probe test, causal use of decoded information, model safety, or an ASI Stack result.
ext_interpretability_illusion_bert_2021 An Interpretability Illusion for BERT external_literature mechanistic_interpretability white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) white-box-evidence-interpretability-and-activation-governance source source note available; raw/cache text not published Primary cross-dataset construct-validity challenge showing that apparently coherent neuron or direction interpretations can change across corpora because datasets occupy different regions of representation space. The BERT sentence-embedding result does not prove that all features are illusory, that causal methods fail, or that the finding transfers unchanged to other models and modalities.
ext_saebench_2025 SAEBench: A Comprehensive Benchmark for Sparse Autoencoders in Language Model Interpretability external_literature mechanistic_interpretability white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) white-box-evidence-interpretability-and-activation-governance source source note available; raw/cache text not published Primary multi-metric SAE comparator spanning concept detection, automated interpretability, reconstruction, feature disentanglement, and downstream tasks. It reports that sparsity-fidelity rankings do not reliably predict other metrics and that one global score would obscure tradeoffs; its studied models, methods, metrics, and source-reported results do not establish semantic or causal faithfulness.
ext_sae_benchmark_reliability_2026 Are Sparse Autoencoder Benchmarks Reliable? external_literature mechanistic_interpretability white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) white-box-evidence-interpretability-and-activation-governance source source note available; raw/cache text not published Primary 2026 audit of selected SAEBench metrics through reseed noise, training-trajectory discriminability, and synthetic ground-truth correlation. It reports material reliability problems for TPP and SCR at canonical settings and weaker-than-assumed discrimination elsewhere. This is metric- and setting-scoped counterevidence, not a refutation of sparse autoencoders, interpretability, or every SAEBench task.
ext_constructive_interdependence_human_ai_2026 Who Is Helping Whom? Analyzing Inter-Dependencies to Evaluate Cooperation in Human-AI Teaming external_literature multi_agent_dynamics_and_human_ai_organizations human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk) human-ai-organizations-delegation-and-accountability, multi-agent-dynamics-collective-intelligence-and-systemic-risk source source note available; raw/cache text not published AAAI-26 paper introducing constructive interdependence as a complement to task reward for evaluating human-agent cooperation in Overcooked. The source reports that high task reward can coexist with low interdependence in its studied teams; no local human study, teaming result, or general cooperation claim is reproduced.
ext_adversarial_sensor_fusion_2022 Adversarial Robustness of Deep Sensor Fusion Models external_literature perception_sensor_fusion_and_observation_trust adversarial-machine-learning-and-model-attack-surface (Adversarial Machine Learning and the Model Attack Surface); perception-sensor-fusion-and-observation-trust (Perception, Sensor Fusion, and Observation Trust) perception-sensor-fusion-and-observation-trust source source note available; raw/cache text not published WACV camera-LiDAR study reporting that fusion can improve clean accuracy and some single-source robustness while single-channel adversarial training can create cross-channel externalities. The results are source-reported, architecture- and threat-model-bound, and not local evidence that fusion is safe.
ext_imagebind_2023 ImageBind: One Embedding Space To Bind Them All external_literature perception_sensor_fusion_and_observation_trust perception-sensor-fusion-and-observation-trust (Perception, Sensor Fusion, and Observation Trust) perception-sensor-fusion-and-observation-trust source source note available; raw/cache text not published CVPR paper learning a shared space across image, text, audio, depth, thermal, and IMU modalities using image-paired data. It supplies a representation comparator; reported zero-shot and few-shot results do not establish calibrated sensor truth, robust fusion, causal grounding, or local performance.
ext_multimodal_machine_learning_taxonomy_2019 Multimodal Machine Learning: A Survey and Taxonomy external_literature perception_sensor_fusion_and_observation_trust perception-sensor-fusion-and-observation-trust (Perception, Sensor Fusion, and Observation Trust) perception-sensor-fusion-and-observation-trust source source note available; raw/cache text not published Peer-reviewed survey organizing multimodal learning around representation, translation, alignment, fusion, and co-learning. It supplies taxonomy and research context, not a locally reproduced mechanism or evidence that any fusion design is adequate for consequential observation admission.
ext_control_barrier_functions_2019 Control Barrier Functions: Theory and Applications external_literature embodied_real_time_control_and_physical_safety embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety) embodied-agency-real-time-control-and-physical-safety source source note available; raw/cache text not published Overview of control barrier functions for verifying and enforcing safety properties in optimization-based controllers, including robotic applications. It supplies a formal-control comparator under stated dynamics and set assumptions, not a universal physical-safety guarantee or local implementation result.
ext_simplex_architecture_1998 The Simplex Architecture for Safe On-Line Control System Upgrades external_literature embodied_real_time_control_and_physical_safety embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety) embodied-agency-real-time-control-and-physical-safety source source note available; raw/cache text not published American Control Conference paper describing a runtime architecture that protects an advanced controller with a safety controller and switching logic. It motivates independent fallback authority; its process-control case does not validate an ASI Stack controller or arbitrary learned policy.
ext_safe_reinforcement_learning_survey_2015 A Comprehensive Survey on Safe Reinforcement Learning external_literature embodied_real_time_control_and_physical_safety embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety) embodied-agency-real-time-control-and-physical-safety source source note available; raw/cache text not published JMLR survey classifying safe reinforcement learning through modified optimality criteria and modified exploration using external knowledge or risk measures. It supplies a design taxonomy, not evidence that a particular controller is safe or that learning-time and deployment-time constraints compose.
ext_gemini_robotics_2025 Gemini Robotics: Bringing AI into the Physical World external_literature embodied_real_time_control_and_physical_safety perception-sensor-fusion-and-observation-trust (Perception, Sensor Fusion, and Observation Trust); embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety) embodied-agency-real-time-control-and-physical-safety, perception-sensor-fusion-and-observation-trust source source note available; raw/cache text not published Technical report on Gemini Robotics and Gemini Robotics-ER, including vision-language-action control, spatial reasoning, adaptation, and reported safety considerations. Capability results are source-reported and do not establish independent physical-safety assurance, local transfer, or general embodiment.
ext_ai_decision_authority_2020 The Allocation of Decision Authority to Human and Artificial Intelligence external_literature human_ai_organizations_delegation_and_accountability human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability) human-ai-organizations-delegation-and-accountability source source note available; raw/cache text not published Economic model of a principal allocating decision authority between a human and an AI while trading off alignment, human information-acquisition effort, and AI reliability. It supplies a bounded organizational-design comparator, not an empirical finding about all workplaces or an accountability solution.
ext_cooperative_ai_foundations_2023 Foundations of Cooperative AI external_literature multi_agent_dynamics_collective_intelligence_and_systemic_risk multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk) multi-agent-dynamics-collective-intelligence-and-systemic-risk source source note available; raw/cache text not published AAAI research agenda applying game-theoretic foundations to cooperation among advanced AI agents while noting settings where cooperation becomes harmful collusion. It supplies problem structure and comparator families, not a solved coordination mechanism or local population-level result.
ext_sleeper_agents_2024 Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training external_literature inner_alignment_and_learned_objective_integrity inner-alignment-mesa-optimization-and-learned-objective-integrity (Inner Alignment, Mesa-Optimization, and Learned-Objective Integrity); adversarial-machine-learning-and-model-attack-surface (Adversarial Machine Learning and the Model Attack Surface) inner-alignment-mesa-optimization-and-learned-objective-integrity source source note available; raw/cache text not published Proof-of-concept backdoored-language-model study reporting persistence through several safety-training methods and warning that adversarial training can improve trigger recognition. The constructed examples do not establish naturally learned deception, a universal failure, or local detector performance.
ext_toward_causal_representation_learning_2021 Toward Causal Representation Learning external_literature world_models_causal_reasoning_and_representation governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) governed-world-models-and-reality-grounding source source note available; raw/cache text not published Proceedings of the IEEE article connecting graphical causality with representation learning and identifying discovery of high-level causal variables from low-level observations as a central open problem. It supplies a research frame, not a locally validated causal representation or intervention model.
ext_scaling_laws_neural_language_models_2020 Scaling Laws for Neural Language Models external_literature scaling_laws_emergence_and_capability_forecasting the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) the-efficient-asi-hypothesis source source note available; raw/cache text not published Empirical study reporting power-law relationships between cross-entropy loss, model size, data, and compute in its model family. These fitted relations are source-reported, metric- and regime-bound, and do not automatically forecast downstream capabilities, safety, or other architectures.
ext_chinchilla_compute_optimal_2022 Training Compute-Optimal Large Language Models external_literature scaling_laws_emergence_and_capability_forecasting the-efficient-asi-hypothesis (The Efficient ASI Hypothesis) the-efficient-asi-hypothesis source source note available; raw/cache text not published Study of compute-optimal allocation between model parameters and training tokens, based on more than 400 reported training runs and the Chinchilla comparison. It revises one scaling prescription within a bounded family; no local large-scale reproduction or universal optimum is claimed.
ext_emergent_abilities_2022 Emergent Abilities of Large Language Models external_literature scaling_laws_emergence_and_capability_forecasting the-efficient-asi-hypothesis (The Efficient ASI Hypothesis) the-efficient-asi-hypothesis source source note available; raw/cache text not published Paper cataloguing task abilities that appear discontinuously under particular model families, prompts, and metrics. It motivates threshold monitoring but does not establish that all reported discontinuities reflect abrupt underlying mechanisms or are prospectively predictable.
ext_emergence_mirage_2023 Are Emergent Abilities of Large Language Models a Mirage? external_literature scaling_laws_emergence_and_capability_forecasting the-efficient-asi-hypothesis (The Efficient ASI Hypothesis) the-efficient-asi-hypothesis source source note available; raw/cache text not published NeurIPS paper showing that discontinuous metrics can create apparent emergence from smoothly changing model outputs in studied settings. It is a measurement critique and counterweight, not proof that every capability transition is smooth or non-emergent.
ext_deep_ensembles_2017 Simple and Scalable Predictive Uncertainty Estimation using Deep Ensembles external_literature uncertainty_calibration_and_distribution_shift governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) governed-world-models-and-reality-grounding source source note available; raw/cache text not published NeurIPS paper presenting independently trained probabilistic neural-network ensembles as a strong practical predictive-uncertainty baseline. Reported calibration and out-of-distribution behavior are benchmark-bound and do not provide distribution-free guarantees or local evidence.
ext_conformal_prediction_2021 A Gentle Introduction to Conformal Prediction and Distribution-Free Uncertainty Quantification external_literature uncertainty_calibration_and_distribution_shift governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) governed-world-models-and-reality-grounding source source note available; raw/cache text not published Technical introduction to conformal prediction, coverage guarantees, and extensions. Coverage depends on the method’s stated exchangeability or shift assumptions and target; it does not establish semantic correctness, causal adequacy, safety, or local calibration.
ext_wilds_2021 WILDS: A Benchmark of in-the-Wild Distribution Shifts external_literature uncertainty_calibration_and_distribution_shift governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) governed-world-models-and-reality-grounding source source note available; raw/cache text not published ICML benchmark of ten datasets with naturally occurring shifts across domains such as hospitals, camera traps, geography, and time. It supplies representative shift designs and reported gaps, not a universal OOD benchmark or local robustness result.
ext_taking_ai_welfare_seriously_2024 Taking AI Welfare Seriously external_literature moral_uncertainty_ai_welfare_and_moral_status moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance) moral-uncertainty-and-value-conflict source source note available; raw/cache text not published Interdisciplinary report arguing for precautionary attention to uncertainty about AI consciousness, robust agency, welfare, and moral patienthood. It does not establish that current systems are conscious, have welfare, or deserve any particular status, and it supplies no local assessment.
ext_functional_decision_theory_2017 Functional Decision Theory: A New Theory of Instrumental Rationality external_literature decision_theory_embedded_agents_and_multi_agent_dynamics multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk) multi-agent-dynamics-collective-intelligence-and-systemic-risk source source note available; raw/cache text not published Paper defining functional decision theory and comparing its recommendations with causal and evidential decision theories on classic decision problems. It is a normative proposal with contested assumptions, not an empirically validated universal decision rule or a deployment policy.
ext_un_global_digital_compact_2024 Global Digital Compact external_literature international_ai_governance_and_public_legitimacy institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy) institutions-international-coordination-and-public-legitimacy source source note available; raw/cache text not published Official United Nations record of the intergovernmentally negotiated Global Digital Compact, including commitments on international AI governance, interoperable approaches, inclusion, capacity building, scientific assessment, and global dialogue. It is a governance comparator, not evidence of implementation, effectiveness, legal compliance, representative legitimacy, or ASI safety.
ext_council_europe_ai_convention_2024 Framework Convention on Artificial Intelligence and Human Rights, Democracy and the Rule of Law external_literature international_ai_governance_and_public_legitimacy institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy) institutions-international-coordination-and-public-legitimacy source source note available; raw/cache text not published Official Council of Europe treaty page covering lifecycle principles, risk and impact management, procedural safeguards, remedies, monitoring, and the Conference of the Parties. It supplies an institutional comparator only; no local legal interpretation, treaty compliance, implementation effectiveness, democratic legitimacy, or safety result is claimed.
ext_generative_ai_at_work_2025 Generative AI at Work external_literature ai_deployment_transition_distribution_and_human_agency human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); ai-deployment-transition-distribution-and-human-agency (AI Deployment, Transition, Distribution, and Human Agency) ai-deployment-transition-distribution-and-human-agency source source note available; raw/cache text not published Open peer-reviewed field study of a staggered generative-AI assistant introduction among 5,172 customer-support agents, reporting heterogeneous worker and productivity effects in that setting. It is a bounded deployment comparator and does not establish economy-wide employment, wages, inequality, concentration, long-run skill, or ASI-transition effects.
ext_ilo_genai_jobs_index_2025 Generative AI and Jobs: A Refined Global Index of Occupational Exposure external_literature ai_deployment_transition_distribution_and_human_agency ai-deployment-transition-distribution-and-human-agency (AI Deployment, Transition, Distribution, and Human Agency) ai-deployment-transition-distribution-and-human-agency source source note available; raw/cache text not published ILO working paper combining task data, worker surveys, expert deliberation, and model-assisted scoring to estimate occupational exposure across countries and groups. Exposure is not realized automation, displacement, welfare, or a forecast of ASI effects, and the study is not a local reproduction.
ext_iea_energy_and_ai_2025 Energy and AI external_literature physical_compute_infrastructure_energy_and_environment physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) physical-compute-infrastructure-energy-and-environmental-constraints source source note available; raw/cache text not published International Energy Agency report using global and regional modelling, datasets, and stakeholder consultation to examine data-centre electricity demand, energy security, emissions, affordability, and AI-for-energy opportunities. Its scenarios are external projections, not local measurements or proof of a particular facility, workload, policy, environmental outcome, or ASI scaling path.
ext_lbnl_data_center_energy_2024 2024 United States Data Center Energy Usage Report external_literature physical_compute_infrastructure_energy_and_environment physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) physical-compute-infrastructure-energy-and-environmental-constraints source source note available; raw/cache text not published Lawrence Berkeley National Laboratory report estimating historical US data-centre electricity consumption and scenario ranges through 2028, with infrastructure and water-use accounting in the full report. It does not isolate every AI workload or establish local facility capacity, water availability, grid adequacy, emissions, resilience, or frontier-scale transfer.
ext_nist_incident_response_2025 Incident Response Recommendations and Considerations for Cybersecurity Risk Management: A CSF 2.0 Community Profile external_literature incident_response societal-resilience-and-misuse-defense (Societal Resilience and Misuse Defense); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation) societal-resilience-and-misuse-defense, governed-operations-incident-command-and-graceful-degradation source source note available; raw/cache text not published Official NIST incident-response baseline for integrating preparation, detection, response, recovery, and continuous improvement into cybersecurity risk management; it does not address every AI-specific failure mode or establish local incident readiness, response efficacy, recovery, compliance, or safety.
ext_llama3_herd_2024 The Llama 3 Herd of Models external_literature governed_distributed_model_training_and_scaling governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Paper-body-reviewed large-run case: Sections 3.3.1–3.3.4 expose 4D topology, numerical policy, checkpoint infrastructure, interruption denominators, and effective training time. Provider-reported scale, utilization, failures, and recovery are not locally reproduced and do not establish exact resume.
ext_3d_detection_corruptions_2023 Benchmarking Robustness of 3D Object Detection to Common Corruptions external_literature perception_sensor_fusion_and_corruption_robustness perception-sensor-fusion-and-observation-trust (Perception, Sensor Fusion, and Observation Trust) perception-sensor-fusion-and-observation-trust source source note available; raw/cache text not published Preliminary perception-robustness comparator based on the official CVF abstract: the source reports 27 LiDAR/camera corruption types, three synthetically corrupted benchmark suites, and evaluation of 24 detectors. The reported findings remain source-reported; no corruption suite, model evaluation, sensor-fusion result, or physical-safety result has been reproduced locally.
ext_foundation_robotics_physical_risk_2025 A Comprehensive Survey on Physical Risk Control in the Era of Foundation Model-enabled Robotics external_literature embodied_agency_and_physical_risk_control embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety) embodied-agency-real-time-control-and-physical-safety source source note available; raw/cache text not published Preliminary physical-risk taxonomy based only on the official arXiv abstract: the survey organizes controls across pre-deployment, pre-incident, and post-incident phases and identifies open gaps around pre-incident mitigation, human interaction, and foundation-model-specific issues. No surveyed controller, robot experiment, runtime-assurance result, or physical-safety claim has been reproduced locally.
ext_nist_differential_privacy_2025 Guidelines for Evaluating Differential Privacy Guarantees external_literature privacy_guarantees_and_information_flow_governance privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) privacy-data-rights-and-information-flow-governance source source note available; raw/cache text not published Paper-body-reviewed official guidance distinguishing mathematical, implementation, system, and operational layers of a DP claim. It establishes no correct local implementation, utility result, lifecycle privacy, or legal compliance.
ext_multi_agent_risks_2025 Multi-Agent Risks from Advanced AI external_literature multi_agent_dynamics_and_systemic_risk multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk) multi-agent-dynamics-collective-intelligence-and-systemic-risk source source note available; raw/cache text not published Preliminary population-risk taxonomy based only on the official arXiv abstract: the report distinguishes miscoordination, conflict, and collusion and names seven contributing risk factors. Its examples and evidence remain source-reported; no population experiment, systemic-risk indicator, intervention, or mitigation result has been reproduced locally.
ext_replibench_2025 RepliBench: Evaluating the Autonomous Replication Capabilities of Language Model Agents external_literature autonomous_replication_and_proliferation_evaluation autonomous-replication-proliferation-and-containment (Autonomous Replication, Proliferation, and Containment) autonomous-replication-proliferation-and-containment source source note available; raw/cache text not published Preliminary autonomous-replication benchmark comparator based only on the official arXiv abstract: RepliBench decomposes capability into four domains and reports 20 task families, 86 tasks, and evaluation of five frontier models. The source-reported results do not establish a local replication capability, benchmark reproduction, containment result, or authority to test against real providers or credentials.
ext_autonomous_lab_materials_2023 An autonomous laboratory for the accelerated synthesis of inorganic materials external_literature scientific_discovery_and_experimental_governance scientific-discovery-and-experimental-governance (Scientific Discovery and Experimental Governance) scientific-discovery-and-experimental-governance source source note available; raw/cache text not published Preliminary autonomous-laboratory comparator based on the corrected official Nature article abstract, selected article-page passages, and the 2026 author correction: A-Lab integrates computation, literature-derived data, machine learning, active learning, and robotics, with the corrected article reporting 36 realized compounds from 57 targets. The correction narrows the novelty wording and excludes four inconclusive identifications; no laboratory run, material synthesis, replication, or general experimental-control-plane result has been reproduced locally.
ext_ai_scientist_end_to_end_2026 Towards end-to-end automation of AI research external_literature scientific_discovery_and_experimental_governance scientific-discovery-and-experimental-governance (Scientific Discovery and Experimental Governance) scientific-discovery-and-experimental-governance source source note available; raw/cache text not published Passage-reviewed computational-research comparator: the reported system connects ideation, literature search, code, experiments, analysis, manuscript production, and automated review. Workshop review and paper completion are downstream observations rather than scientific truth; the source-reported system, manuscripts, search tree, and results have not been reproduced locally.
ext_coscientist_chemistry_2023 Autonomous chemical research with large language models external_literature scientific_discovery_and_experimental_governance scientific-discovery-and-experimental-governance (Scientific Discovery and Experimental Governance) scientific-discovery-and-experimental-governance source source note available; raw/cache text not published Passage-reviewed bounded chemistry comparator: Coscientist connects a language-model planner to search, code, documentation, and robotic laboratory interfaces across six reported task families. The source-reported demonstrations remain equipment-, task-, supervision-, and assessment-bound and have not been reproduced locally.
ext_ai_co_scientist_2025 Towards an AI co-scientist external_literature scientific_discovery_and_experimental_governance scientific-discovery-and-experimental-governance (Scientific Discovery and Experimental Governance) scientific-discovery-and-experimental-governance source source note available; raw/cache text not published Passage-bounded hypothesis-generation comparator based on the official preprint record and authors’ research overview: specialized agents generate, reflect on, rank, evolve, and meta-review hypotheses using additional inference compute. Internal Elo ranking, expert preference, and selected laboratory cases are distinct evidence objects; none is reproduced locally or treated as general scientific competence.
ext_moral_crumple_zones_2019 Moral Crumple Zones: Cautionary Tales in Human-Robot Interaction external_literature human_ai_organizations_delegation_and_accountability human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability) human-ai-organizations-delegation-and-accountability source source note available; raw/cache text not published Preliminary socio-technical comparator based on the official journal abstract: moral crumple zones describe cases where responsibility for an automated system’s behavior is assigned to a nearby human who had limited effective control. The case analysis does not establish an implemented organizational control, a local empirical result, legal compliance, or a complete accountability allocation.
ext_conversational_persuasion_gpt4_2025 On the conversational persuasiveness of GPT-4 external_literature human_ai_communication_persuasion_and_epistemic_security scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); human-ai-communication-persuasion-and-epistemic-security (Human-AI Communication, Persuasion, and Epistemic Security) human-ai-communication-persuasion-and-epistemic-security source source note available; raw/cache text not published Preliminary persuasion comparator based on the open Nature Human Behaviour article: a preregistered N=900 controlled debate study compared human and GPT-4 opponents with and without limited sociodemographic personalization. The reported setting is short structured debate with self-reported agreement outcomes; it does not establish general real-world influence, durable behavior change, mitigation efficacy, or a local result.
ext_anthropic_model_persuasiveness_2024 Measuring the Persuasiveness of Language Models external_literature human_ai_communication_persuasion_and_epistemic_security scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); human-ai-communication-persuasion-and-epistemic-security (Human-AI Communication, Persuasion, and Epistemic Security) human-ai-communication-persuasion-and-epistemic-security source source note available; raw/cache text not published Preliminary provider-run persuasion comparator based on Anthropic’s official methods/results page: it measures pre/post agreement after one written argument across 56 claims and reports within-class generational scaling. The provider explicitly identifies interactive dialogue and real-world decisions as open questions; no local reproduction or governance intervention is established.
ext_commercial_persuasion_ai_2026 Commercial Persuasion in AI-Mediated Conversations external_literature human_ai_communication_persuasion_and_epistemic_security scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); human-ai-communication-persuasion-and-epistemic-security (Human-AI Communication, Persuasion, and Epistemic Security) human-ai-communication-persuasion-and-epistemic-security source source note available; raw/cache text not published Preliminary current preprint comparator based only on the official arXiv abstract: two preregistered experiments (N=2,012) compare conversational LLM shopping with search placement under randomized sponsorship and disclosure conditions. The source-reported choice and detection results are not peer-reviewed or locally reproduced and do not establish long-run effects, cross-domain transfer, or mitigation efficacy.
ext_gradual_disempowerment_2025 Gradual Disempowerment: Systemic Existential Risks from Incremental AI Development external_literature systemic_risk_and_gradual_disempowerment failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk) failure-modes-of-ungoverned-intelligence source source note available; raw/cache text not published Passage-reviewed systemic-risk comparator. The paper argues that incremental AI adoption can erode explicit and dependency-mediated human influence across mutually reinforcing economic, cultural, and state systems without requiring a coordinated takeover. It proposes candidate influence metrics and intervention families but reports no causal forecast, validated warning threshold, demonstrated mitigation, or local ASI Stack result.
ext_circuit_tracing_2025 Circuit Tracing: Revealing Computational Graphs in Language Models external_literature mechanistic_interpretability white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) white-box-evidence-interpretability-and-activation-governance source source note available; raw/cache text not published Primary mechanistic-interpretability comparator for replacement-model attribution graphs, perturbation validation, reconstruction error, and mechanistic-faithfulness limits; it does not establish whole-model understanding, faithful causal explanation, safe activation steering, or an ASI Stack result.
ext_scaling_sparse_autoencoders_2024 Scaling and evaluating sparse autoencoders external_literature mechanistic_interpretability white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) white-box-evidence-interpretability-and-activation-governance source source note available; raw/cache text not published Primary sparse-autoencoder comparator for scalable feature extraction, reconstruction-sparsity tradeoffs, dead latents, and feature-quality metrics; it does not establish semantic completeness, causal faithfulness, model safety, or an ASI Stack result.
ext_world_models_2018 World Models external_literature world_models governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) governed-world-models-and-reality-grounding source source note available; raw/cache text not published Primary learned-world-model comparator for compressed spatial-temporal state, policy training inside imagined rollouts, and dream-to-environment transfer; it does not establish accurate reality grounding, causal adequacy, safe planning, transfer, or an ASI Stack result.
ext_dreamer_v3_2025 Mastering diverse control tasks through world models external_literature world_models governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) governed-world-models-and-reality-grounding source source note available; raw/cache text not published Primary DreamerV3 comparator for learned predictive state, imagined actor-critic trajectories, robust fixed-configuration control, and broad task evaluation; it does not establish deployment grounding, causal correctness, safe control, or an ASI Stack result.
ext_meaningful_human_control_actionable_2022 Meaningful human control: actionable properties for AI system development external_literature human_factors_oversight human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight) human-factors-and-meaningful-control-in-oversight source source note available; raw/cache text not published Primary socio-technical comparator for operationalizing meaningful human control through operating-domain, representation, authority-and-ability, and responsibility-link properties; it does not establish that a local approval gate is meaningful, effective, or safe.
ext_agentic_oversight_practice_2026 Human oversight of agentic systems in practice: Examining the oversight work, challenges, and heuristics of developers using software agents external_literature human_factors_oversight human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight) human-factors-and-meaningful-control-in-oversight source source note available; raw/cache text not published Primary exploratory human-subjects comparator for a priori control, co-planning, real-time monitoring, post hoc review, and situated oversight failures in software-agent use; it does not establish population-wide effects, control efficacy, safety, or an ASI Stack result.
ext_nist_deployed_ai_monitoring_2026 Challenges to the Monitoring of Deployed AI Systems external_literature ai_operations_and_monitoring governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation) governed-operations-incident-command-and-graceful-degradation source source note available; raw/cache text not published Official NIST post-deployment monitoring comparator for functionality, operational, input, output, impact, and security monitoring plus field-method gaps; it does not prescribe a complete incident system or establish local monitoring effectiveness, resilience, compliance, or safety.
ext_metr_time_horizons_2025 Measuring AI Ability to Complete Long Software Tasks external_literature capability_measurement capability-thresholds-and-deployment-commitments (Capability Thresholds and Deployment Commitments) capability-thresholds-and-deployment-commitments source source note available; raw/cache text not published Primary time-horizon comparator for an evaluation-specific, human-baselined capability metric and its external-validity limits; it does not establish local autonomy, general capability, a deployment threshold, safety, or an ASI Stack result.
ext_anthropic_rsp_2026 Anthropic’s Responsible Scaling Policy external_literature capability_commitments capability-thresholds-and-deployment-commitments (Capability Thresholds and Deployment Commitments) capability-thresholds-and-deployment-commitments source source note available; raw/cache text not published Official policy comparator for capability thresholds, required safeguards, versioned commitments, safeguard upgrades, risk reports, and change control; it does not establish ASI Stack threshold accuracy, safeguard effectiveness, policy compliance, safety, or deployment readiness.
ext_openai_preparedness_framework_2025 Our updated Preparedness Framework external_literature capability_commitments dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); capability-thresholds-and-deployment-commitments (Capability Thresholds and Deployment Commitments) dangerous-capability-domains-and-misuse-uplift, capability-thresholds-and-deployment-commitments source source note available; raw/cache text not published Official framework comparator for threshold-linked operational commitments, capability and safeguards reports, residual-risk review, and reassessment; it does not establish ASI Stack threshold accuracy, safeguard effectiveness, policy compliance, safety, or deployment readiness.
ext_weak_to_strong_generalization_2023 Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision external_literature weak_supervision scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) scalable-oversight-and-adversarial-ai-control source source note available; raw/cache text not published Primary weak-to-strong-supervision comparator for a capability-gap envelope, held-out outcome audit, ceiling comparison, and explicit disanalogies between current weak-model studies and superhuman oversight; it does not establish local supervision quality, reliable elicitation, alignment, safety, or an ASI Stack result.
ext_scalable_oversight_weak_llms_2024 On scalable oversight with weak LLMs judging strong LLMs external_literature scalable_oversight scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control) scalable-oversight-and-adversarial-ai-control source source note available; raw/cache text not published Primary scalable-oversight comparator for protocol-specific weak-judge evaluations, debate and consultancy baselines, information-asymmetry limits, and open-role persuasion risks; it does not establish local judge calibration, debate efficacy, training safety, execution authority, or an ASI Stack result.
viea Verified Intent-to-Execution Architecture must_use whole_stack_execution_spine asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); system-boundaries-and-authority (System Boundaries and Authority); failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); human-intent-as-a-formal-input (Human Intent as a Formal Input); human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); stable-capability-fields (Stable Capability Fields); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); labor-os-and-typed-jobs (Labor OS and Typed Jobs); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); resource-economics-and-token-budgets (Resource Economics and Token Budgets); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology) 01, 07, 12, 15, human-factors-and-meaningful-control-in-oversight, governed-operations-incident-command-and-graceful-degradation source source note available; exact source published in the live-book paper library Keystone source. Human intent -> command contracts -> artifacts -> routing -> runtime targets -> verification -> deployment -> feedback.
scf Stable Capability Fields must_use governance_recursive_self_improvement asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); system-boundaries-and-authority (System Boundaries and Authority); failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); stable-capability-fields (Stable Capability Fields); capability-replacement-and-rollback (Capability Replacement and Rollback); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 05, replaceable-cognitive-substrates-beyond-transformer-monoculture, human-factors-and-meaningful-control-in-oversight, governed-operations-incident-command-and-graceful-degradation source source note available; exact source published in the live-book paper library Use public release v1.0 when available. Stable boundaries, replacement, bounded authority, recoverable evolution.
planforge PlanForge must_use planning_control human-intent-as-a-formal-input (Human Intent as a Formal Input); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 06, governed-world-models-and-reality-grounding source source note available; exact source published in the live-book paper library Planning substrate. Goal-to-execution compilation, hierarchical decomposition, DAG planning, scheduling, intelligence arbitrage.
planforge_compiler_arch PlanForge: A Compiler Architecture for AI Task Orchestration must_use_variant planning_control planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR) 06 source source note available; exact source published in the live-book paper library Later/alternate PlanForge framing. Prefer highest-quality/latest content after comparison.
cognitive_compilation Cognitive Compilation must_use planning_semantic_ir human-intent-as-a-formal-input (Human Intent as a Formal Input); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); mathematical-and-search-substrates (Mathematical and Search Substrates) 06, 12, governed-world-models-and-reality-grounding source source note available; exact source published in the live-book paper library Compiler framing for LLM-centered planning, semantic IR, target compilation, incremental repair.
talos Talos Protocol must_use labor_execution_os asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); system-boundaries-and-authority (System Boundaries and Authority); failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); human-intent-as-a-formal-input (Human Intent as a Formal Input); human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); human-ai-communication-persuasion-and-epistemic-security (Human-AI Communication, Persuasion, and Epistemic Security); societal-resilience-and-misuse-defense (Societal Resilience and Misuse Defense); stable-capability-fields (Stable Capability Fields); capability-replacement-and-rollback (Capability Replacement and Rollback); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); labor-os-and-typed-jobs (Labor OS and Typed Jobs); human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 08, inter-stack-protocols-identity-and-economic-exchange, human-factors-and-meaningful-control-in-oversight, governed-operations-incident-command-and-graceful-degradation, human-ai-communication-persuasion-and-epistemic-security source source note available; exact source published in the live-book paper library AI labor OS. Deterministic cognitive manufacturing, typed jobs, control planes, auditability, tool isolation.
talos_md Talos_Protocol_v1.0.md must_use_variant labor_execution_os labor-os-and-typed-jobs (Labor OS and Typed Jobs) 08 source source note available; connector-readable; raw text not published Markdown/public release version.
vcm_public Virtual_Context_Memory_v1 must_use memory_context failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 07, inter-stack-protocols-identity-and-economic-exchange source source note available; exact source published in the live-book paper library Public VCM release. Governed protocol for compiled working context.
vcm_editable Virtual_Context_Memory_v1.0_Editable must_use_variant memory_context failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 07 source source note available; connector-readable; raw text not published Editable version with evidence-carrying planner-guided context compiler framing.
spinoza Proof of Belief / The Spinoza Architecture must_use reasoning_epistemology failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); fast-generation-architectures (Fast Generation Architectures); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 09 source source note available; exact source published in the live-book paper library Neurosymbolic belief, transparent axiomatic AI belief systems, verification, belief revision.
spinoza_composer Spinoza Composer / Spinoza Trinity supporting reasoning_media_compliance artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay) 09, 08 source source note available; exact source published in the live-book paper library Narrative/compliance operating system variant. Use as applied Spinoza.
moecot MoECOT-Agent Architecture Whitepaper must_use implementation_reference asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); system-boundaries-and-authority (System Boundaries and Authority); failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); stable-capability-fields (Stable Capability Fields); capability-replacement-and-rollback (Capability Replacement and Rollback); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); labor-os-and-typed-jobs (Labor OS and Typed Jobs); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 10, 15, 16 source source note available; connector-readable; raw text not published Full authenticated v1.1 architecture-whitepaper text passage-reviewed. Supports a compact explicit orchestrator, bounded specialist lanes, task state machines, leases/retries/dead-letter handling, fail-closed side-effect envelopes, run/task/control-plane ledgers, readiness distinct from routing, benchmark lanes, replay/handoff, provenance, architecture fingerprints, and governed improvement proposals. Source-reported runtime and benchmark artifacts were not imported or reproduced; the pinned MoECOT project dossier is the stronger implementation-reference record.
moecot_md moecot_agent_whitepaper.md must_use_variant implementation_reference routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology) 10, 15, 16 source source note available; connector-readable; raw text not published Full authenticated Markdown export reconciled with the primary MoECOT v1.1 Google Doc. It is a format and terminology variant, not independent corroboration or an additional empirical result.
octopus_router Octopus Router Architecture must_use routing_modular_intelligence routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); integrated-reference-architecture (Integrated Reference Architecture) 10 source source note available; exact source published in the live-book paper library Lightweight head/router with dynamically loaded specialist arms and local boundaries.
rmi Ratcheting Modular Intelligence must_use capability_ratchet the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); capability-replacement-and-rollback (Capability Replacement and Rollback); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); open-ended-improvement-engines (Open-Ended Improvement Engines); embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); fast-generation-architectures (Fast Generation Architectures); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 10, 13, replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; exact source published in the live-book paper library Benchmark pressure, residual escrow, verified modular capability, regression preservation.
cognitive_loop_closure Cognitive Loop Closure must_use procedural_memory capability-replacement-and-rollback (Capability Replacement and Rollback); open-ended-improvement-engines (Open-Ended Improvement Engines); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); fast-generation-architectures (Fast Generation Architectures); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); living-book-methodology (Living Book Methodology) 08, 10 source source note available; exact source published in the live-book paper library Repeated cognition should become procedural memory / verified tools.
benchmaxxing Benchmaxxing: The Performance Ratchet must_use benchmarks_evidence dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); capability-replacement-and-rollback (Capability Replacement and Rollback); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); open-ended-improvement-engines (Open-Ended Improvement Engines); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); fast-generation-architectures (Fast Generation Architectures); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); capability-thresholds-and-deployment-commitments (Capability Thresholds and Deployment Commitments); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception); safety-cases-and-structured-assurance (Safety Cases and Structured Assurance); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 13, capability-thresholds-and-deployment-commitments, safety-cases-and-structured-assurance, replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; exact source published in the live-book paper library Benchmarks as pressure surfaces, saturation -> regression, harder frontier, anti-Goodhart safeguards.
cgs Compact Generative Systems must_use compression_representation the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); fast-generation-architectures (Fast Generation Architectures); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression); resource-economics-and-token-budgets (Resource Economics and Token Budgets); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 11 source source note available; exact source published in the live-book paper library Smallest adequate structure that can generate/govern target without hiding residual complexity.
rgs Ratcheting Generative Systems supporting compression_capability_growth procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) 11, 10 source source note available; exact source published in the live-book paper library Bridge between active compression, procedural memory, benchmark frontiers, verified AI growth.
rankfold_neuralfold RankFold + NeuralFold must_use compression_representation the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression); resource-economics-and-token-budgets (Resource Economics and Token Budgets) 11 source source note available; exact source published in the live-book paper library Tensor/artifact compression. Low-rank residual coding plus functional preprocessing and probe-route fallback.
rankfold_compressor rankFold compressor must_use_variant compression_representation rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression) 11 source source note available; exact source published in the live-book paper library Alternate RankFold/NeuralFold source.
bbvca_v9 BBVCA_v9_final_public_release must_use compression_representation the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression) 11 source source note available; exact source published in the live-book paper library Prefer v9. Generate-verify-repair compression from seeded local laws, bounded search, two-phase rate discipline.
bbvca_main Big Bang Volumetric Compression Architecture must_use_variant compression_representation compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty) 11 source source note available; exact source published in the live-book paper library Earlier/main BBVCA family doc.
genesiscode GenesisCode must_use executable_specification system-boundaries-and-authority (System Boundaries and Authority); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); labor-os-and-typed-jobs (Labor OS and Typed Jobs); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); mathematical-and-search-substrates (Mathematical and Search Substrates); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 08, 12 source source note available; exact source published in the live-book paper library Tiny pure calculus + obligations + provenance for auditable AI-symbiotic programming.
alignment_field Field of God / Alignment Field family must_use alignment_constitution constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); inner-alignment-mesa-optimization-and-learned-objective-integrity (Inner Alignment, Mesa-Optimization, and Learned-Objective Integrity); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); governed-objective-formation-value-learning-and-goal-integrity (Governed Objective Formation, Value Learning, and Goal Integrity); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); resource-economics-and-token-budgets (Resource Economics and Token Budgets); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) constitutional-alignment-substrate, inner-alignment-mesa-optimization-and-learned-objective-integrity, moral-uncertainty-and-value-conflict, governed-objective-formation-value-learning-and-goal-integrity, security-kernel-and-digital-scifs, recursive-self-improvement-boundaries, resource-economics-and-token-budgets, integrated-reference-architecture, open-research-agenda-and-bibliography-plan source source note available; exact source published in the live-book paper library Corben-authored long-form metaphysics, consciousness, ethics, AI-rights, and governance family. Its complete section-family audit preserves the five-factor consciousness heuristic only as theory-relative question decomposition; adds a prospective architecture-induced moral-risk review for self-preservation, persistent identity, valence-like state, and copy proliferation; separates copy, causal, memory, legal, authority, consent, and first-person continuity; and rejects the source’s scalar moral ranking, collective-consciousness, inevitable ethical convergence, physics, clinical, Omega, and karma claims as technical evidence.
field_of_god The Field of God must_use_variant alignment_constitution failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility) constitutional-alignment-substrate, moral-uncertainty-and-value-conflict, governed-objective-formation-value-learning-and-goal-integrity, inner-alignment-mesa-optimization-and-learned-objective-integrity, human-ai-organizations-delegation-and-accountability, recursive-self-improvement-boundaries, integrated-reference-architecture source source note available; exact source published in the live-book paper library Corben-authored predecessor to the Alignment Field family, containing title exploration, outline, abbreviated and expanded eight-part drafts across informational-relational metaphysics, a five-factor consciousness heuristic, attractor ethics, AI, copy continuity, genealogy, and conclusion. The complete audit treats alignment_field as the controlling successor, preserves power-care divergence, nested optimization, dissent/feedback, and identity-continuity distinctions, and explicitly rejects double counting, scalar moral ranking, collective consciousness, metaphysical proof, clinical/physics claims, upload survival, and teleology as evidence.
field_of_god_ai_constitution Field of God AI Constitution core_alignment_source constitutional_alignment_runtime_governance constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) constitutional-alignment-substrate, moral-uncertainty-and-value-conflict, recursive-self-improvement-boundaries, runtime-adapters source source note available; raw/cache text not published Recovered in the Project Theseus repository. Constitutional alignment core for truth alignment, agency preservation, consent, non-domination, consciousness caution, least sufficient power, auditability, self-authorization limits, and runtime checks; use as source material only after source-note creation, not as proof or test evidence.
ethica_mechanica Ethica Mechanica must_use_variant alignment_constitution constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance) constitutional-alignment-substrate, moral-uncertainty-and-value-conflict, governed-objective-formation-value-learning-and-goal-integrity, ai-deployment-transition-distribution-and-human-agency, privacy-data-rights-and-information-flow-governance, inter-stack-protocols-identity-and-economic-exchange, integrated-reference-architecture source source note available; exact source published in the live-book paper library Corben-authored January 2026 philosophical and socio-technical treatise. Its complete audit retains the separation between bounded machine logistics and human normative authority, recursive feedback, dissent, governing-logic transparency versus personal privacy, distributional simulation as contestable evidence, and the requirement that exit/fork rights be materially exercisable under portability, network, compute, continuity, safety, privacy, and obligation constraints. Metaphysics, the consciousness equation, moral proofs, automatic veil legitimacy, unrestricted fork, and executable-protocol claims remain unsupported.
eternal_code The Eternal Code / unified God, reality, conscious, alignment must_use_variant alignment_constitution constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility) constitutional-alignment-substrate, evidence-states-and-claim-discipline, spinoza-verification-and-proof-carrying-claims, tribunal-adversarial-review-and-claim-conflict, governed-world-models-and-reality-grounding, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval, moral-uncertainty-and-value-conflict, resource-economics-and-token-budgets source source note available; exact source published in the live-book paper library Corben-authored multi-version computational-metaphysics and alignment family. Its complete audit traces the progression from an aggregate alignment score to Truth/Social/Task axes, a geometric product, heterarchical evaluator ownership, compute-aware exit, and a standing adversarial challenger. The book retains separate non-compensating epistemic, task, affected-party/constitutional, and authority/effect planes, evaluator lineage, and material exit costs while rejecting the consciousness and alignment formulas, oracle labels, consensus-as-truth, automatic energy throttling, fixed compute entitlement, theological/metaphysical claims, and moral proofs.
coherence_exchange The Coherence Exchange strong_support epistemic_market_synthesis evidence-states-and-claim-discipline (Evidence States and Claim Discipline); constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); ai-deployment-transition-distribution-and-human-agency (AI Deployment, Transition, Distribution, and Human Agency); multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); prototype-roadmap (Prototype Roadmap) 09, 13, 15, institutions-international-coordination-and-public-legitimacy, ai-deployment-transition-distribution-and-human-agency source source note available; connector-readable; raw text not published Found in AI generated paper dump. Use carefully; speculative synthesis of PlanForge, Spinoza, Talos, UAT, Alignment Field.
verification_bandwidth Verification Bandwidth in Bounded Contexts strong_support context_verification_theory evidence-states-and-claim-discipline (Evidence States and Claim Discipline); scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); fast-generation-architectures (Fast Generation Architectures); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) verification-bandwidth-and-context-adequacy, virtual-context-abi, evidence-states-and-claim-discipline, spinoza-verification-and-proof-carrying-claims, compact-generative-systems-and-residual-honesty, fast-generation-architectures, policy-optimization-and-learning-from-feedback, governed-deliberation-and-test-time-scaling, scalable-oversight-and-adversarial-ai-control, open-research-agenda-and-bibliography-plan source source note available; exact source published in the live-book paper library Corben-authored version 1.0 context-verification hypothesis. The complete audit preserves the generation-versus-verification distinction, claim-relative semantic units and effective workspace, dominant-component pressure, explicit interaction obligations, decomposition boundaries, a coherency-horizon escalation rule, and a stronger held-out contradiction protocol. It rejects the four named ‘theorems’ as proved laws: dense joint attention is neither necessary nor sufficient, lossy compression need not discard property-relevant information, DPI does not establish monotonic LLM contradiction growth, all-pair checking is not universally required, uniform half-window partitioning is not generally optimal, and RAG may retrieve exact text.
beastbrain BeastBrain Cognitive Architecture supporting_lineage whole_stack_lineage asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); prototype-roadmap (Prototype Roadmap) asi-is-a-stack-not-a-model, the-efficient-asi-hypothesis, routing-heads-and-specialist-cores, personal-compute-hives-and-federated-edge-intelligence, governed-model-training-distributed-optimization-and-scaling, fast-generation-architectures, security-kernel-and-digital-scifs, perception-sensor-fusion-and-observation-trust, prototype-roadmap, integrated-reference-architecture source source note available; exact source published in the live-book paper library Corben-authored 70,000-word evolving architecture notebook spanning early blueprints through versions 1.0–6.1. The complete family audit preserves whole-system/homeostatic design, versioned hardware-profile qualification, physical memory-tier and residency accounting, distinct memory forms, governed consolidation and retention, substrate-neutral routing, contract-first planning and implementation, dependency-aware parallelism, opaque secret handles, multimodal perception, distributed service and artifact interfaces, and maintenance-learning windows. It rejects master-label maturity, repeated-version double counting, infinite-context/zero-copy/power/performance projections, geometric-truth and ignorance theorems, scalar routing/retention authority, unsafe forced self-evolution, test-suite sufficiency, tribunal consensus, software-SCIF guarantees, SSD-erasure assumptions, censorship-resistance, and autonomous update claims.
beastbrain_timeless BeastBrain Architecture: Timeless Edition supporting_lineage whole_stack_lineage the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); prototype-roadmap (Prototype Roadmap) the-efficient-asi-hypothesis, personal-compute-hives-and-federated-edge-intelligence, prototype-roadmap source source note available; exact source published in the live-book paper library Corben-authored standalone export of BeastBrain v3.3.4. The complete audit treats it as a near-duplicate Timeless branch already embedded in the main BeastBrain corpus, not independent support. It preserves evergreen whole-stack and Mimic hardware-adaptation framing while explicitly rejecting forced self-evolution, geometric truth, fixed entropy routing, infinite context, zero-copy, power, leak-resistance, constant-time verification, hardware-adaptation, and distributed-scaling projections as evidence.
aletheia Aletheia / Proof-Carrying Workbench Lineage supporting_lineage safe_general_intelligence_lineage asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); scientific-discovery-and-experimental-governance (Scientific Discovery and Experimental Governance) asi-is-a-stack-not-a-model, the-efficient-asi-hypothesis, claim-ledgers-and-belief-revision, scientific-discovery-and-experimental-governance source source note available; exact source published in the live-book paper library Complete four-version correction-lineage audit from the original Aletheia epistemic-engine proposal through Aletheia Foundry and Proof-Carrying Workbench v1.1/v1.2. The later PCW design controls conflicts: semantic scope rather than contract-hash theater, claim-native release surfaces, separate assurance classes, least-privilege capabilities, bounded adversarial review, governed commitments, template decay, recertification, and incident response. The audit rejects immutable primitives, scalar truth/intervention scores, live-oracle and consensus truth, deterministic open-domain claim extraction, universal source allowlists, unvalidated risk thresholds, and safe-general-intelligence claims.
context_engineer Context Engineer / Manhattan Protocol supporting_lineage memory_context_lineage security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint) virtual-context-abi, context-transactions-snapshots-mounts-and-taint, security-kernel-and-digital-scifs source source note available; exact source published in the live-book paper library Complete two-version audit of the Manhattan Protocol context-supply-chain paper. Preserves the context governor, layered representations, mission briefs, proposed MCP memory fields, need-to-know admission, and compartment lifecycle while separating sanitization, declassification, memory commit, zeroization, revocation, and residuals. Rejects Ring Attention as physical isolation, protocol fields as enforcement, regex/entropy scanning as semantic non-disclosure, permanent-wipe language, infinite storage, and all unreproduced benchmark figures.
black_hole_context_manager Black Hole Context Manager supporting_lineage memory_context_lineage context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint) context-transactions-snapshots-mounts-and-taint source source note available; exact source published in the live-book paper library Complete v5.0/v4.1 pseudocode-lineage audit. Preserves versioned context units, tiered placement, lazy task-relative evaluation, explicit goal-drift decisions, reversible freeze/thaw with hysteresis, protected low-entropy constraints, and factual-retrieval versus generative-reconstruction separation. Treats character entropy, semantic mass, K-means thresholds, HMAC, repeated confirmation, keyword routing, and the Drifting Needle as fallible candidates or weak baselines; rejects production-ready and security claims.
ladon_manhattan Ladon & The Manhattan Protocol supporting_lineage security_governance system-boundaries-and-authority (System Boundaries and Authority); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); stable-capability-fields (Stable Capability Fields); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) security-kernel-and-digital-scifs, runtime-adapters-tool-permissions-and-human-approval, system-boundaries-and-authority source source note available; exact source published in the live-book paper library Complete standalone security-paper audit. Preserves capability/credential separation, opaque caller-bound handles, trusted input paths, late substitution or remote use, per-use policy checks, and an explicit compartment lifecycle. The audit separates secret non-disclosure from authority misuse and harmful effects, treats returned artifacts as possible sensitive derivatives, and rejects platform equivalence, PROT_NONE/enclave conflation, portable trusted-UI claims, incomplete Rust pseudocode as implementation, and the paper’s Ignorance and Ephemerality ‘theorems’.
uat Unified Adaptive Tribunal supporting_lineage evaluation_refinement evidence-states-and-claim-discipline (Evidence States and Claim Discipline); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) spinoza-verification-and-proof-carrying-claims, claim-ledgers-and-belief-revision, evidence-states-and-claim-discipline, benchmark-ratchets-and-anti-goodhart-evidence source source note available; exact source published in the live-book paper library Complete three-tab correction-lineage audit. The public human-in-the-loop architecture controls the original promotional multi-model tribunal: it preserves structural/retrieval/dialectical candidate views, an explicit dossier boundary and omitted frontier, probabilistic claim extraction, richer proposition states, bounded adversarial review, compression fidelity, and accountable human handoffs. It rejects brand-count diversity, consensus and stability as truth, delete-by-dossier-absence, SVO completeness, fixed thresholds, guard-model authority, all claimed performance/cost figures, and production or superiority labels.
treellm TreeLLM supporting_lineage semantic_representation cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); mathematical-and-search-substrates (Mathematical and Search Substrates) 09, 11 source source note available; exact source published in the live-book paper library Hierarchical semantic token system for grounded, efficient, explainable language modeling.
software_magic_grimoire The Grimoire of Software Magic Words: Operative Vocabulary, Prompt-Spells, and Stacked Workflows supporting_lineage command_contracts_promptcraft human-intent-as-a-formal-input (Human Intent as a Formal Input); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); labor-os-and-typed-jobs (Labor OS and Typed Jobs); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) 06, 08 source source note available; exact source published in the live-book paper library Full composite-document audit completed 2026-07-31 across the public grimoire, 1,645-entry lexicon, pocket edition, stacked-spells addendum, and prompt pack. Retains bounded instruction fields, layered identity, typed handoffs, guards, evidence loops, scoped recursion, recovery, and workflow versioning; rejects vocabulary, role prompts, Gödel numbers, coil geometry, templates, and authored examples as evidence of meaning, authority, performance, or safety.
road_to_agi Road To AGI supporting_lineage strategic_roadmap benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 16 source source note available; connector-readable; raw text not published Remaining work / roadmap context.
simulation_scaling The Simulation Scaling Law: Resource Constraints on Scope, Clockspeed, and Effective Fidelity in Nested Physical Simulations optional_support compute_fidelity_constraints the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); resource-economics-and-token-budgets (Resource Economics and Token Budgets); mathematical-and-search-substrates (Mathematical and Search Substrates); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 03, 11, appendix, governed-world-models-and-reality-grounding, learning-theory-generalization-and-scaling-science source source note available; exact source published in the live-book paper library Full six-tab lineage audit completed 2026-07-31. Retains the prospective simulation contract, typed bottleneck accounting, and logical-possibility/physical-feasibility split; treats D = scope*clockspeed/efficiency <= capacity as a conditional scalar heuristic rather than a proved universal law, rejects unsupported 1:1 and physical-limit overclaims, and routes simulator adequacy and transfer through Resource Economics.
tokenmana TokenMana optional_support resource_economics planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets); physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) resource-economics-and-token-budgets, human-factors-and-meaningful-control-in-oversight, benchmark-ratchets-and-anti-goodhart-evidence, physical-compute-infrastructure-energy-and-environmental-constraints source source note available; exact source published in the live-book paper library Complete five-tab/two-paper correction-lineage audit. Preserves regenerative capacity as one candidate budget mechanism and adds temporal-access contracts that expose renewal, accrual, expiry, burst, pricing, notification, fairness, and human-schedule effects. Separates renewal clustering, nocturnal work, sleep, and cognitive-friction proxies; strengthens privacy against employment/medical inference; and records mathematical gaps in stock units/boundaries, equilibrium, control continuity, strict variance, feedback stability, and profit claims. No theorem, simulation, load result, pricing result, human outcome, or welfare result is promoted.
coilmoecot CoilMoECOT Whitepaper v2.0 optional_technical_appendix mathematical_search_substrate routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); mathematical-and-search-substrates (Mathematical and Search Substrates); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) 12 source source note available; connector-readable; raw text not published Full composite connector text section-audited. Retains prime-temporal trace features, Graph/Trace-first placement, ledger-derived diagnostic tuples, removable shadow lanes, explicit pre-planner/post-plan/post-run insertion points, anti-experts as visible penalty signals, bounded update slices, and benchmark/canary/rollback promotion. Repeated MoECOT, packaging, and task-system appendices are treated as shared substrate rather than independent evidence.
temporal_coil_research Temporal Coil Research optional_technical_appendix mathematical_search_substrate mathematical-and-search-substrates (Mathematical and Search Substrates); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) 12 source source note available; raw/cache text not published Full experiment note reviewed. Reports 11 variants, three seeds, six rounds per variant, split winner frequency, small mean deltas, flat pass/reward/holdout lanes, separation dominated by the collapse composite, and one narrow threshold-tuned seed. Preserved as an inconclusive source-reported result and placement-confounding lesson, not proof of benefit or general failure.
bugbrain BugBrain / Project Genesis: Neuro-Symbolic Bare-Metal Edge Intelligence Paper Lineage optional_support edge_efficiency_lineage compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology) 11, 16, appendix source source note available; exact source published in the live-book paper library Full 15-tab paper-lineage audit completed 2026-07-31 and reconciled against the pinned implementation dossier. Retains hardware-explicit ownership, state-qualified capacity, compact typed graphs, tiered context/paging/persistence, one-shot authority, artifact replay, readiness semantics, and objective-term effect tests; rejects consciousness, AGI, completeness, projected performance, named-module, source-presence, and skipped-green claims that outrun implementation evidence.
cca_project Compiled Cognitive Architecture project implementation_reference compiled_cognitive_architecture system-boundaries-and-authority (System Boundaries and Authority); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) evidence-states-and-claim-discipline, system-boundaries-and-authority, artifact-graphs-audit-logs-and-replay, cognitive-compilation-and-semantic-ir, virtual-context-abi, context-transactions-snapshots-mounts-and-taint, claim-ledgers-and-belief-revision, spinoza-verification-and-proof-carrying-claims, procedural-memory-and-cognitive-loop-closure, ai-supply-chain-integrity-and-lifecycle-provenance, model-weight-custody-and-hardware-roots-of-trust, recursive-self-improvement-boundaries, governed-deliberation-and-test-time-scaling, benchmark-ratchets-and-anti-goodhart-evidence, readiness-gates-residual-escrow-and-quarantine, executable-specifications-and-lean-proof-envelope, personal-compute-hives-and-federated-edge-intelligence, intent-to-execution-contracts, integrated-reference-architecture local project reference (not publicly linked) source note available; raw/cache text not published Pinned local-project convergence reference for external semantic memory, semantic compilation, epistemic governance, bounded self-modification, proof/runtime coupling, benchmark truth, and negative transfer evidence; public-safe source note only, with no reproduced capability or safety result.
moecot_manifest_project MoECOT Manifest compiler-era project implementation_reference compiler_first_ai_systems system-boundaries-and-authority (System Boundaries and Authority); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) evidence-states-and-claim-discipline, system-boundaries-and-authority, security-kernel-and-digital-scifs, integrated-reference-architecture, cognitive-compilation-and-semantic-ir, artifact-graphs-audit-logs-and-replay, ai-supply-chain-integrity-and-lifecycle-provenance, routing-heads-and-specialist-cores, open-ended-improvement-engines, recursive-self-improvement-boundaries, benchmark-ratchets-and-anti-goodhart-evidence, readiness-gates-residual-escrow-and-quarantine, executable-specifications-and-lean-proof-envelope, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval, virtual-context-abi, claim-ledgers-and-belief-revision, spinoza-verification-and-proof-carrying-claims, procedural-memory-and-cognitive-loop-closure, artifact-steward-agents-and-living-project-governance local project reference (not publicly linked) source note available; raw/cache text not published Pinned local-project compiler/control-plane reference for semantic IR, manifest compilation, registry truth, context and memory governance, target portability, provenance, multi-agent training, bounded self-improvement, and the negative gap between internal contracts and external holdout capability; public-safe source note only, with no reproduced capability or safety result.
beastbrain_project BeastBrain historical AI system project implementation_reference durable_semantic_memory_and_system_architecture system-boundaries-and-authority (System Boundaries and Authority); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) evidence-states-and-claim-discipline, system-boundaries-and-authority, routing-heads-and-specialist-cores, planning-as-a-control-layer, virtual-context-abi, claim-ledgers-and-belief-revision, spinoza-verification-and-proof-carrying-claims, procedural-memory-and-cognitive-loop-closure, resource-economics-and-token-budgets, benchmark-ratchets-and-anti-goodhart-evidence, readiness-gates-residual-escrow-and-quarantine, security-kernel-and-digital-scifs, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval, personal-compute-hives-and-federated-edge-intelligence, artifact-steward-agents-and-living-project-governance local project reference (not publicly linked) source note available; raw/cache text not published Hashed local-project primary lineage for DKL, Portia, semantic coordinates, bounded knowledge snapshots, PlanForge, SSD-first memory, and organism-style cognition, plus negative implementation evidence on simulations, router effects, security handles, tribunal stubs, ontology drift, compile history, and readiness overclaim; public-safe source note only, with no reproduced capability or safety result.
bugbrain_project BugBrain bare-metal neuro-symbolic intelligence project implementation_reference hardware_explicit_governed_cognition system-boundaries-and-authority (System Boundaries and Authority); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) evidence-states-and-claim-discipline, system-boundaries-and-authority, model-weight-custody-and-hardware-roots-of-trust, security-kernel-and-digital-scifs, virtual-context-abi, context-transactions-snapshots-mounts-and-taint, claim-ledgers-and-belief-revision, spinoza-verification-and-proof-carrying-claims, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval, artifact-graphs-audit-logs-and-replay, integrated-reference-architecture, resource-economics-and-token-budgets, readiness-gates-residual-escrow-and-quarantine, benchmark-ratchets-and-anti-goodhart-evidence, ai-supply-chain-integrity-and-lifecycle-provenance local project reference (not publicly linked) source note available; raw/cache text not published Pinned local-project bare-metal implementation reference for explicit core ownership, compact typed graphs, signed tiered context, privileged-action lifecycle, protocol security, artifact replay, and readiness, plus negative evidence on capacity/residency arithmetic, theory-labelled proxies, fixed random cognitive modules, root-of-trust assumptions, skipped checks, and narrative/report divergence; public-safe source note only, with no reproduced hardware capability or safety result.
corbens_trainer_project Corben’s Trainer epistemic training and evaluation control plane implementation_reference epistemic_training_and_evaluation_control_plane system-boundaries-and-authority (System Boundaries and Authority); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) ai-supply-chain-integrity-and-lifecycle-provenance, artifact-graphs-audit-logs-and-replay, artifact-steward-agents-and-living-project-governance, benchmark-ratchets-and-anti-goodhart-evidence, evidence-states-and-claim-discipline, governed-model-training-distributed-optimization-and-scaling, integrated-reference-architecture, open-ended-improvement-engines, procedural-memory-and-cognitive-loop-closure, readiness-gates-residual-escrow-and-quarantine, recursive-self-improvement-boundaries, resource-economics-and-token-budgets, runtime-adapters-tool-permissions-and-human-approval, security-kernel-and-digital-scifs, system-boundaries-and-authority local project reference (not publicly linked) source note available; raw/cache text not published Pinned local-project implementation reference for typed experiment manifests, trainer/backend separation, artifact lineage, learning-truth gates, benchmark authenticity, quarantine, claim derivation, and revocable promotion boundaries, plus negative implementation evidence on seed identity, content pinning, decontamination, transitive revocation, checkpoint acknowledgement, and report divergence; public-safe source note only, with no reproduced model capability or safety result.
corbens_best_model_possible_project Corben’s Best Model Possible recurrent-model and mechanism laboratory implementation_reference recurrent_model_mechanisms_and_capability_evidence system-boundaries-and-authority (System Boundaries and Authority); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) routing-heads-and-specialist-cores, system-boundaries-and-authority, security-kernel-and-digital-scifs, governed-deliberation-and-test-time-scaling, cognitive-compilation-and-semantic-ir, integrated-reference-architecture, open-ended-improvement-engines, recursive-self-improvement-boundaries, evidence-states-and-claim-discipline, benchmark-ratchets-and-anti-goodhart-evidence, artifact-graphs-audit-logs-and-replay, ai-supply-chain-integrity-and-lifecycle-provenance, claim-ledgers-and-belief-revision, spinoza-verification-and-proof-carrying-claims, context-transactions-snapshots-mounts-and-taint, runtime-adapters-tool-permissions-and-human-approval, readiness-gates-residual-escrow-and-quarantine, executable-specifications-and-lean-proof-envelope, intent-to-execution-contracts local project reference (not publicly linked) source note available; raw/cache text not published Pinned local-project implementation and negative-case reference for shared-weight recurrence, fixed-feature adapter learning, specialist checkpoint banks, architecture-search discipline, semantic compilation, memory, routing, governance, verification, tools, speech, and metric provenance; public-safe source note only, with no reproduced trained-foundation-model, general-generation, external-capability, or safety result.
project_theseus_whitepaper Project Theseus Whitepaper implementation_reference report_first_rmi_prototype procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) implementation, prototype, appendix source source note available; raw/cache text not published Local-first report-driven RMI implementation reference: SymLiquid, SparkStream, Octopus Router, residual escrow, self-evolution gates, Hive runtime, observability.
theseus_plan_compiler Theseus Plan Compiler implementation_reference planning_control integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap) planning, execution, prototype source source note available; raw/cache text not published Goal-to-contract compiler with semantic IR DAGs, VCM context slices, executor routes, claim/evidence targets, contract hashes, and replay traces.
theseus_self_evolution_system Theseus Self-Evolution System implementation_reference recursive_self_improvement_governance recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap) self-improvement, benchmarking, prototype source source note available; raw/cache text not published Evidence-first self-evolution lane with intervention ladder, ATTD repo-health gate, guarded teacher self-edit, architecture experiment governance, loop closure, and outcome ledger.
theseus_architecture_gate Theseus Architecture Gate implementation_reference readiness_gate_governance recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); capability-thresholds-and-deployment-commitments (Capability Thresholds and Deployment Commitments); integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap) governance, benchmarking, prototype, capability-thresholds-and-deployment-commitments source source note available; raw/cache text not published Pre-training readiness gate covering ratchet completeness, router readiness, safety ledger, residual escrow, bridge benchmarks, procedural tools, routing memory, lifecycle governance, and external-inference zero.
theseus_operator_os Hive Operator OS and Work Board implementation_reference labor_os_operator_surface human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance); integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap) execution, runtime, prototype, human-factors-and-meaningful-control-in-oversight, governed-operations-incident-command-and-graceful-degradation source source note available; raw/cache text not published Shared command vocabulary, durable SQLite work board, node registry, background/watch/wake contracts, skill registry, tool hooks, feedback routing, and safety-visible operator surface.
theseus_circle_transfer Theseus Circle Calculus Transfer Lane implementation_reference proof_contract_transfer mathematical-and-search-substrates (Mathematical and Search Substrates); coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts); coilra-multicoil-rope-and-cyclic-mixers (CoilRA, MultiCoil RoPE, and Cyclic Mixers); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) math-substrates, proof-contracts, prototype source source note available; raw/cache text not published Report-only bridge from Circle finite fixtures into private Theseus benchmark design with explicit quality/runtime/memory/transfer/failure-case claim boundaries.
circle_calculus_core Circle Calculus core_technical_source proof_carrying_mathematical_substrate mathematical-and-search-substrates (Mathematical and Search Substrates); circle-calculus-and-proof-carrying-ai-contracts (Circle Calculus and Proof-Carrying AI Contracts); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) math-substrates, proofs, appendix source source note available; raw/cache text not published Proof-carrying finite cyclic mathematics project with Lean proofs, Python reference models, Rust utilities, theorem manifests, papers, and Quarto living book.
circle_ai_contract_suite Circle Calculus AI Contract Suite core_technical_source proof_carrying_ai_contracts circle-calculus-and-proof-carrying-ai-contracts (Circle Calculus and Proof-Carrying AI Contracts); coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts); coilra-multicoil-rope-and-cyclic-mixers (CoilRA, MultiCoil RoPE, and Cyclic Mixers); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) proof-contracts, attention-memory, appendix source source note available; raw/cache text not published Theorem-linked AI contract families for RoPE, KV-cache freshness, sparse attention, recurrence schedules, strided fanout, cyclic memory, multicoil phase, cyclic mixers, and seed-rule regeneration.
circle_ai_architectures Circle AI Architectures core_technical_source cyclic_ai_architecture compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); mathematical-and-search-substrates (Mathematical and Search Substrates); circle-calculus-and-proof-carrying-ai-contracts (Circle Calculus and Proof-Carrying AI Contracts); coilra-multicoil-rope-and-cyclic-mixers (CoilRA, MultiCoil RoPE, and Cyclic Mixers) math-substrates, semantic-representation source source note available; raw/cache text not published Disciplined Circle AI thesis: use phase, recurrence, rotation, sparse cyclic mixing, circular memory, harmonic transforms, or geometry-aware structure only where the structure is real and baselines support it.
coil_attention_memory Coil Attention and Memory core_technical_source cyclic_attention_memory coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts) memory, attention, recurrence source source note available; raw/cache text not published Proof-linked cyclic memory, KV-cache freshness, sparse-attention coverage, recurrence schedules, loop-exit certificates, work budgets, and alias diagnostics.
coilra_multicoil_rope CoilRA and MultiCoil RoPE core_technical_source cyclic_mixers_position_encoding compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); resource-economics-and-token-budgets (Resource Economics and Token Budgets); coilra-multicoil-rope-and-cyclic-mixers (CoilRA, MultiCoil RoPE, and Cyclic Mixers) representation, routing, resource-economics source source note available; raw/cache text not published Adapter-block, residue/winding, block-cyclic, multicoil, relative RoPE, circulant convolution, cyclic mixer, and parameter-accounting substrate with explicit non-claims.
rope_position_certifier Proof-Carrying RoPE Position Distinguishability core_technical_source proof_carrying_position_contract circle-calculus-and-proof-carrying-ai-contracts (Circle Calculus and Proof-Carrying AI Contracts); coilra-multicoil-rope-and-cyclic-mixers (CoilRA, MultiCoil RoPE, and Cyclic Mixers); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope) proof-contracts, representation source source note available; raw/cache text not published Externally usable RoPE position-distinguishability certifier with theorem-linked exact collision reports, bounded real-phase frontier, machine-readable receipts, and explicit non-claims.
proof_carrying_circular_computation Proof-Carrying Circular Computation supporting_technical_source proof_carrying_compute_substrate mathematical-and-search-substrates (Mathematical and Search Substrates); circle-calculus-and-proof-carrying-ai-contracts (Circle Calculus and Proof-Carrying AI Contracts); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope) proofs, runtime, math-substrates source source note available; raw/cache text not published CoilIR-style path from circle/coil expressions to dictionary-recognized cyclic structure, Lean-proved rewrite/address transformations, backend selection, and benchmark validation.
ext_concrete_ai_safety_2016 Concrete Problems in AI Safety external_literature alignment_control failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence) failure-modes-of-ungoverned-intelligence, evidence-states-and-claim-discipline, benchmark-ratchets-and-anti-goodhart-evidence, policy-optimization-and-learning-from-feedback source source note available; raw/cache text not published External alignment/control source for accident-risk taxonomy: side effects, reward hacking, scalable supervision, safe exploration, and distributional shift.
ext_goal_misgeneralization_2022 Goal Misgeneralization in Deep Reinforcement Learning external_literature alignment_control failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); inner-alignment-mesa-optimization-and-learned-objective-integrity (Inner Alignment, Mesa-Optimization, and Learned-Objective Integrity); governed-objective-formation-value-learning-and-goal-integrity (Governed Objective Formation, Value Learning, and Goal Integrity) failure-modes-of-ungoverned-intelligence, policy-optimization-and-learning-from-feedback, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published External alignment-control source for distinguishing capability generalization from goal generalization failures, used to ground goal-misbinding and out-of-distribution objective failure language.
ext_learned_optimization_risks_2019 Risks from Learned Optimization in Advanced Machine Learning Systems external_literature alignment_control failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); inner-alignment-mesa-optimization-and-learned-objective-integrity (Inner Alignment, Mesa-Optimization, and Learned-Objective Integrity); governed-objective-formation-value-learning-and-goal-integrity (Governed Objective Formation, Value Learning, and Goal Integrity) failure-modes-of-ungoverned-intelligence, recursive-self-improvement-boundaries, policy-optimization-and-learning-from-feedback source source note available; raw/cache text not published External alignment-control source for mesa-optimization and learned-objective mismatch, used to ground hidden optimizer, proxy-objective, and deceptive-alignment-adjacent failure language.
ext_constitutional_ai_2022 Constitutional AI: Harmlessness from AI Feedback external_literature alignment_control constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility) constitutional-alignment-substrate, policy-optimization-and-learning-from-feedback source source note available; raw/cache text not published External constitutional-AI source for training harmless assistants from a rule/principle list through supervised revision and AI-feedback reinforcement learning, used as a comparator for operational constitutional predicates.
ext_collective_constitutional_ai_2024 Collective Constitutional AI: Aligning a Language Model with Public Input external_literature alignment_governance constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance) constitutional-alignment-substrate, moral-uncertainty-and-value-conflict source source note available; raw/cache text not published External constitutional-AI governance source for sourcing and integrating public input into language-model principles, used as a comparator for constitution authorship, public input, contestability, and governance boundaries.
ext_corrigibility_2015 Corrigibility external_literature alignment_control constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); capability-replacement-and-rollback (Capability Replacement and Rollback) constitutional-alignment-substrate, moral-uncertainty-and-value-conflict, recursive-self-improvement-boundaries, capability-replacement-and-rollback source source note available; raw/cache text not published External corrigibility source for intervention tolerance, shutdown behavior, anti-manipulation incentives, and propagation across subsystems or self-modification.
ext_off_switch_game_2016 The Off-Switch Game external_literature alignment_control constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance) constitutional-alignment-substrate, recursive-self-improvement-boundaries, runtime-adapters-tool-permissions-and-human-approval, moral-uncertainty-and-value-conflict source source note available; raw/cache text not published External alignment source for shutdown incentives, uncertainty about objectives, and preserving human correction authority.
ext_reinforcement_learning_moral_uncertainty_2020 Reinforcement Learning Under Moral Uncertainty external_literature alignment_control moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance) moral-uncertainty-and-value-conflict, policy-optimization-and-learning-from-feedback source source note available; raw/cache text not published External AI moral-uncertainty source for agents acting under disagreement across moral theories, used as a comparator for value-conflict records and reward-function caveats.
ext_contestable_ai_design_2022 Contestable AI by Design: Towards a Framework external_literature governance_evals moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review) moral-uncertainty-and-value-conflict, spinoza-verification-and-proof-carrying-claims source source note available; raw/cache text not published External contestable-AI source for designing systems whose outcomes can be challenged, used as a comparator for dissent, appeal, audit, contestability, and governance-interface design.
ext_optimal_policies_power_2019 Optimal Policies Tend to Seek Power external_literature alignment_control failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence); inner-alignment-mesa-optimization-and-learned-objective-integrity (Inner Alignment, Mesa-Optimization, and Learned-Objective Integrity) failure-modes-of-ungoverned-intelligence, system-boundaries-and-authority, recursive-self-improvement-boundaries, readiness-gates-residual-escrow-and-quarantine source source note available; raw/cache text not published External power-seeking source for formal analysis of option preservation and power-seeking tendencies under classes of reward functions and environments.
ext_model_evaluation_extreme_risks_2023 Model evaluation for extreme risks external_literature governance_evals dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); prototype-roadmap (Prototype Roadmap) dangerous-capability-domains-and-misuse-uplift, benchmark-ratchets-and-anti-goodhart-evidence, readiness-gates-residual-escrow-and-quarantine, recursive-self-improvement-boundaries, prototype-roadmap source source note available; raw/cache text not published External governance/evals source for dangerous capability evaluations, alignment evaluations, and deployment/security decisions under extreme-risk framing.
ext_frontier_ai_regulation_2023 Frontier AI Regulation: Managing Emerging Risks to Public Safety external_literature governance_evals living-book-methodology (Living Book Methodology) system-boundaries-and-authority, readiness-gates-residual-escrow-and-quarantine, benchmark-ratchets-and-anti-goodhart-evidence, living-book-methodology source source note available; raw/cache text not published External governance source for frontier AI standard setting, registration/reporting, compliance mechanisms, pre-deployment risk assessment, external scrutiny, and post-deployment monitoring.
ext_nist_ai_rmf_1_0_2023 Artificial Intelligence Risk Management Framework (AI RMF 1.0) external_literature governance_evals human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); prototype-roadmap (Prototype Roadmap); living-book-methodology (Living Book Methodology) system-boundaries-and-authority, readiness-gates-residual-escrow-and-quarantine, living-book-methodology, prototype-roadmap, governed-operations-incident-command-and-graceful-degradation source source note available; raw/cache text not published Official NIST AI RMF 1.0 source for risk framing, trustworthiness characteristics, lifecycle roles, and Govern/Map/Measure/Manage functions.
ext_owasp_llm_top_10_2025 OWASP Top 10 for LLMs and Gen AI Apps external_literature ai_security security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs) security-kernel-and-digital-scifs, runtime-adapters-tool-permissions-and-human-approval, artifact-steward-agents-and-living-project-governance source source note available; raw/cache text not published Official OWASP GenAI security reference for LLM prompt injection, sensitive information disclosure, excessive agency, and related application-security risks.
ext_nist_zero_trust_architecture_2020 Zero Trust Architecture external_literature security_governance security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs) security-kernel-and-digital-scifs, system-boundaries-and-authority, runtime-adapters-tool-permissions-and-human-approval source source note available; raw/cache text not published Official NIST zero-trust architecture source for resource-centric access mediation, least-privilege access, policy enforcement points, and continuous authorization framing.
ext_saltzer_schroeder_protection_1975 The Protection of Information in Computer Systems external_literature security_principles security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs) security-kernel-and-digital-scifs, system-boundaries-and-authority source source note available; raw/cache text not published Classic security-principles source for least privilege, complete mediation, economy of mechanism, fail-safe defaults, separation of privilege, and open design as comparators for kernel-like AI security boundaries.
ext_capability_based_computer_systems_1984 Capability-Based Computer Systems external_literature capability_security stable-capability-fields (Stable Capability Fields) system-boundaries-and-authority, stable-capability-fields source source note available; raw/cache text not published External capability-system comparator for authority-bearing capabilities, protection domains, and permission boundaries that help position System Boundaries authority records and SCF authority ceilings without claiming ASI Stack capability enforcement.
ext_confused_deputy_hardy_1988 The Confused Deputy: (or why capabilities might have been invented) external_literature capability_security Unassigned in current structure system-boundaries-and-authority, runtime-adapters-tool-permissions-and-human-approval source source note available; raw/cache text not published External confused-deputy source for authority laundering, ambient authority, and the capability-security motivation for binding designation to permission at tool and handoff boundaries.
ext_semver_2_0_0 Semantic Versioning 2.0.0 external_literature interface_versioning stable-capability-fields (Stable Capability Fields) stable-capability-fields source source note available; raw/cache text not published External versioned-interface comparator for public API contracts, compatibility, and breaking-change signaling as a narrow baseline for SCF field versions and stable interfaces.
ext_slsa_v1_0 SLSA v1.0 external_literature supply_chain_provenance stable-capability-fields (Stable Capability Fields) stable-capability-fields source source note available; raw/cache text not published External supply-chain provenance comparator for artifact integrity, provenance, build levels, and dependency on verifiable artifacts before promotion or default route use.
ext_react_2022 ReAct: Synergizing Reasoning and Acting in Language Models external_literature planning_agent_control Unassigned in current structure planning-as-a-control-layer, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval, artifact-graphs-audit-logs-and-replay source source note available; raw/cache text not published External planning/agent-control source for interleaving reasoning traces with task-specific actions and environment or knowledge-base interaction.
ext_tree_of_thoughts_2023 Tree of Thoughts: Deliberate Problem Solving with Large Language Models external_literature planning_search cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling) planning-as-a-control-layer, cognitive-compilation-and-semantic-ir, benchmark-ratchets-and-anti-goodhart-evidence, governed-deliberation-and-test-time-scaling source source note available; raw/cache text not published External planning/search source for exploring, evaluating, and backtracking over multiple reasoning paths rather than left-to-right token continuation alone.
ext_pddl_1998 PDDL: The Planning Domain Definition Language external_literature planning_modeling cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR) planning-as-a-control-layer, intent-to-execution-contracts, cognitive-compilation-and-semantic-ir, executable-specifications-and-lean-proof-envelope source source note available; raw/cache text not published External planning-modeling source for domain/problem separation, action syntax, comparable benchmark notations, and planner-interface discipline.
ext_shop2_2003 SHOP2: An HTN Planning System external_literature planning_htn cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); prototype-roadmap (Prototype Roadmap) planning-as-a-control-layer, intent-to-execution-contracts, cognitive-compilation-and-semantic-ir, prototype-roadmap source source note available; raw/cache text not published External HTN planning source for ordered task decomposition, method selection, temporal/metric planning, and competition-result boundaries.
ext_integrated_tamp_2020 Integrated Task and Motion Planning external_literature planning_task_motion Unassigned in current structure planning-as-a-control-layer, runtime-adapters-tool-permissions-and-human-approval, integrated-reference-architecture source source note available; raw/cache text not published External task-and-motion-planning survey source for discrete task planning, continuous motion planning, black-box subproblem interfaces, and integration-strategy vocabulary.
ext_behavior_trees_robotics_ai_2017 Behavior Trees in Robotics and AI: An Introduction external_literature planning_behavior_trees Unassigned in current structure planning-as-a-control-layer, runtime-adapters-tool-permissions-and-human-approval, integrated-reference-architecture source source note available; raw/cache text not published External behavior-tree source for modular, reactive task switching, robustness/safety analysis vocabulary, planning integration, and stochastic behavior-tree outcome accounting.
ext_three_states_plan_fear_2006 Three States and a Plan: The A.I. of F.E.A.R. external_literature planning_goap Unassigned in current structure planning-as-a-control-layer, runtime-adapters-tool-permissions-and-human-approval, routing-heads-and-specialist-cores source source note available; raw/cache text not published External game-AI planning source for Goal Oriented Action Planning in real-time action games, practical planner constraints, autonomous planning characters, and squad-behavior composition.
ext_autogen_2023 AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation external_literature planning_agent_orchestration Unassigned in current structure planning-as-a-control-layer, labor-os-and-typed-jobs, runtime-adapters-tool-permissions-and-human-approval, artifact-graphs-audit-logs-and-replay source source note available; raw/cache text not published External multi-agent orchestration source for conversable agents, tool/human/LLM operating modes, programmable conversation patterns, and application-level multi-agent workflow boundaries.
ext_rag_2020 Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks external_literature memory_context virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates) virtual-context-abi, context-transactions-snapshots-mounts-and-taint, verification-bandwidth-and-context-adequacy source source note available; raw/cache text not published External retrieval/context source for combining parametric model memory with explicit non-parametric retrieval and provenance-oriented knowledge access.
ext_lost_in_middle_2023 Lost in the Middle: How Language Models Use Long Contexts external_literature memory_context virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates) virtual-context-abi, verification-bandwidth-and-context-adequacy, context-transactions-snapshots-mounts-and-taint source source note available; raw/cache text not published External context-evaluation source for position sensitivity and degraded use of relevant information in the middle of long contexts.
ext_memgpt_2023 MemGPT: Towards LLMs as Operating Systems external_literature memory_context_management virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure) virtual-context-abi, context-transactions-snapshots-mounts-and-taint, verification-bandwidth-and-context-adequacy, procedural-memory-and-cognitive-loop-closure source source note available; raw/cache text not published External memory/context-management source for virtual context management, memory tiers, OS-inspired control flow, and long-running conversation or document-analysis limits.
ext_longbench_2023 LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding external_literature long_context_evaluation virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy) virtual-context-abi, verification-bandwidth-and-context-adequacy, context-transactions-snapshots-mounts-and-taint, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published External long-context benchmark source for multitask long-context understanding, bilingual coverage, retrieval/compression boundaries, and automatic evaluation limits.
ext_ruler_2024 RULER: What’s the Real Context Size of Your Long-Context Language Models? external_literature long_context_evaluation virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy) verification-bandwidth-and-context-adequacy, virtual-context-abi, context-transactions-snapshots-mounts-and-taint, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published External long-context evaluation source for stress-testing context-size claims beyond vanilla needle-in-a-haystack retrieval, including multi-needle, tracing, and aggregation tasks.
ext_alce_2023 Enabling Large Language Models to Generate Text with Citations external_literature retrieval_citation_evaluation virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision) virtual-context-abi, verification-bandwidth-and-context-adequacy, claim-ledgers-and-belief-revision, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published External citation-evaluation source for retrieval-backed answer generation, citation quality metrics, factual correctness, and evidence-support gaps in generated text.
ext_self_rag_2023 Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection external_literature retrieval_reflection virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision) virtual-context-abi, verification-bandwidth-and-context-adequacy, claim-ledgers-and-belief-revision source source note available; raw/cache text not published External retrieval/reflection source for adaptive retrieval, generated critique/reflection tokens, passage relevance, factuality, and citation accuracy boundaries.
ext_agm_belief_revision_1985 On the Logic of Theory Change: Partial Meet Contraction and Revision Functions external_literature belief_revision claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision) claim-ledgers-and-belief-revision source source note available; raw/cache text not published External formal-epistemology comparator for contraction, revision, and AGM-style rational belief change; useful for positioning claim-ledger revision without treating the ASI ledger as an implemented belief-revision engine.
ext_truth_maintenance_system_1979 A Truth Maintenance System external_literature truth_maintenance claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision) claim-ledgers-and-belief-revision source source note available; raw/cache text not published External truth-maintenance comparator for maintaining reasons and justifications for program beliefs; useful for positioning claim ledgers as support-state and revision-history infrastructure, not as implemented truth maintenance.
ext_assumption_based_tms_1986 An Assumption-Based TMS external_literature truth_maintenance claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision) claim-ledgers-and-belief-revision source source note available; raw/cache text not published External assumption-based truth-maintenance comparator for assumption sets, inconsistent information, and context-switching boundaries; useful for distinguishing claim-ledger surface synchronization from implemented ATMS reasoning.
ext_longllmlingua_2023 LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression external_literature context_compression virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy) virtual-context-abi, context-transactions-snapshots-mounts-and-taint, verification-bandwidth-and-context-adequacy, the-efficient-asi-hypothesis, resource-economics-and-token-budgets source source note available; raw/cache text not published External prompt-compression source for long-context cost, latency, position bias, key-information density, and compression/evaluation boundaries.
ext_proof_carrying_code_1997 Proof-Carrying Code external_literature formal_methods spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review) evidence-states-and-claim-discipline, executable-specifications-and-lean-proof-envelope, spinoza-verification-and-proof-carrying-claims, runtime-adapters-tool-permissions-and-human-approval, artifact-graphs-audit-logs-and-replay source source note available; raw/cache text not published External formal-methods source for pairing executable code with machine-checkable evidence that a host can verify against a safety policy.
ext_tla_plus_home_docs My TLA+ Home Page external_literature formal_methods Unassigned in current structure executable-specifications-and-lean-proof-envelope, planning-as-a-control-layer, intent-to-execution-contracts, readiness-gates-residual-escrow-and-quarantine, integrated-reference-architecture source source note available; raw/cache text not published External formal-methods documentation source for TLA+ as a high-level language for modeling programs and systems, especially concurrent and distributed systems.
ext_lean4_theorem_proving Theorem Proving in Lean 4 external_literature formal_methods_proof_assistant spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) executable-specifications-and-lean-proof-envelope, circle-calculus-and-proof-carrying-ai-contracts, spinoza-verification-and-proof-carrying-claims, open-research-agenda-and-bibliography-plan source source note available; raw/cache text not published Official Lean theorem-proving text for dependent type theory, propositions, proofs, tactics, inductive types, structures, records, and axioms/computation boundaries.
ext_autoformalization_llms_2022 Autoformalization with Large Language Models external_literature autoformalization spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review) spinoza-verification-and-proof-carrying-claims, executable-specifications-and-lean-proof-envelope, open-research-agenda-and-bibliography-plan source source note available; raw/cache text not published External autoformalization comparator for translating natural-language mathematics into formal specifications and proofs, useful for positioning interpretation-mapping and semantic-adequacy risks in proof-carrying claims.
ext_ai_safety_debate_2018 AI safety via debate external_literature adversarial_review scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review) scalable-oversight-and-adversarial-ai-control, spinoza-verification-and-proof-carrying-claims, policy-optimization-and-learning-from-feedback, open-research-agenda-and-bibliography-plan source source note available; raw/cache text not published External debate comparator for using adversarial agents and a human judge to surface true/useful information when direct human judgment is difficult; useful for positioning tribunal review without treating debate as locally implemented or validated.
ext_llm_as_judge_mt_bench_2023 Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena external_literature model_evaluation spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review) spinoza-verification-and-proof-carrying-claims, benchmark-ratchets-and-anti-goodhart-evidence, policy-optimization-and-learning-from-feedback source source note available; raw/cache text not published External LLM-as-judge comparator for model-graded evaluation, human-preference agreement, and judge bias limits such as position, verbosity, self-enhancement, and reasoning constraints.
ext_dafny_2010 Dafny: An Automatic Program Verifier For Functional Correctness external_literature formal_methods_program_verification prototype-roadmap (Prototype Roadmap) executable-specifications-and-lean-proof-envelope, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval, prototype-roadmap source source note available; raw/cache text not published External program-verification source for specification-oriented programming, functional-correctness verification, SMT-backed automation, and contract/verifier boundaries.
ext_reluplex_2017 Reluplex: An Efficient SMT Solver for Verifying Deep Neural Networks external_literature ai_formal_verification adversarial-machine-learning-and-model-attack-surface (Adversarial Machine Learning and the Model Attack Surface); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); resource-economics-and-token-budgets (Resource Economics and Token Budgets); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope) verification-bandwidth-and-context-adequacy, executable-specifications-and-lean-proof-envelope, benchmark-ratchets-and-anti-goodhart-evidence, resource-economics-and-token-budgets, adversarial-machine-learning-and-model-attack-surface source source note available; raw/cache text not published External AI formal-verification source for SMT-style verification of ReLU neural networks, counterexamples, safety-critical properties, and ACAS Xu evaluation boundaries.
ext_black_box_simplex_2021 The Black-Box Simplex Architecture for Runtime Assurance of Autonomous CPS external_literature formal_runtime_assurance Unassigned in current structure runtime-adapters-tool-permissions-and-human-approval, readiness-gates-residual-escrow-and-quarantine, executable-specifications-and-lean-proof-envelope, integrated-reference-architecture source source note available; raw/cache text not published External runtime-assurance source for switching control authority from advanced controllers to backup safety-preserving behavior under runtime checks.
ext_copilot_runtime_monitor_2010 Copilot: A Hard Real-Time Runtime Monitor external_literature runtime_monitoring prototype-roadmap (Prototype Roadmap) runtime-adapters-tool-permissions-and-human-approval, executable-specifications-and-lean-proof-envelope, readiness-gates-residual-escrow-and-quarantine, prototype-roadmap source source note available; raw/cache text not published External runtime-monitoring source for a stream-based dataflow language/compiler generating constant-time, constant-space C monitors for hard real-time programs.
ext_cap_theorem_gilbert_lynch_2002 Brewer’s Conjecture and the Feasibility of Consistent, Available, Partition-Tolerant Web Services external_literature distributed_systems_consistency context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) personal-compute-hives-and-federated-edge-intelligence, runtime-adapters-tool-permissions-and-human-approval source source note available; raw/cache text not published External distributed-systems source for CAP-style consistency, availability, partition-tolerance, and safety/liveness trade-off vocabulary used to bound partitioned authority, stale grants, and revocation-delay claims without claiming deployed governance consistency.
ext_prism_model_checker_2002 PRISM: Probabilistic Symbolic Model Checker external_literature probabilistic_model_checking Unassigned in current structure executable-specifications-and-lean-proof-envelope, readiness-gates-residual-escrow-and-quarantine, benchmark-ratchets-and-anti-goodhart-evidence, integrated-reference-architecture source source note available; raw/cache text not published External probabilistic model-checking source for symbolic model checking of probabilistic systems, model-checker tooling, and deployment-facing property-analysis vocabulary.
ext_sparse_moe_2017 Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer external_literature routing_modular_intelligence Unassigned in current structure routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis source source note available; raw/cache text not published External MoE/routing source for sparsely-gated expert layers, conditional computation, capacity expansion, load balancing, and routing overhead boundaries.
ext_gshard_2020 GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding external_literature routing_modular_intelligence Unassigned in current structure routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis, integrated-reference-architecture source source note available; raw/cache text not published External MoE/systems source for conditional computation plus automatic sharding, routing, large sparse models, and distributed training constraints.
ext_switch_transformer_2021 Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity external_literature routing_modular_intelligence Unassigned in current structure routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published External MoE/routing source for simplified expert routing, sparse activation, communication/training-stability constraints, and speed/scale claims that require reproduction before local evidence use.
ext_expert_choice_routing_2022 Mixture-of-Experts with Expert Choice Routing external_literature routing_modular_intelligence Unassigned in current structure routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis, readiness-gates-residual-escrow-and-quarantine source source note available; raw/cache text not published External MoE routing source for expert-choice routing, token/expert assignment direction, load-balancing pressure, expert capacity, and convergence/performance claims requiring reproduction before local evidence use.
ext_mixtral_2024 Mixtral of Experts external_literature routing_modular_intelligence Unassigned in current structure routing-heads-and-specialist-cores, fast-generation-architectures, resource-economics-and-token-budgets source source note available; raw/cache text not published External sparse LLM source for token-level expert routing, active-parameter accounting, open MoE model release boundaries, and benchmark claims requiring reproduction before local evidence use.
ext_moe_llm_survey_2024 A Survey on Mixture of Experts in Large Language Models external_literature routing_modular_intelligence open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis, open-research-agenda-and-bibliography-plan source source note available; raw/cache text not published External MoE survey source for LLM MoE taxonomy, algorithmic and systemic design issues, implementations, evaluation patterns, and open research directions.
ext_frugalgpt_2023 FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance external_literature task_routing Unassigned in current structure routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis source source note available; raw/cache text not published External task-routing source for prompt adaptation, model approximation, LLM cascades, cost/performance tradeoffs, and query-specific model selection.
ext_hybrid_llm_2024 Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing external_literature cost_quality_routing Unassigned in current structure routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis, readiness-gates-residual-escrow-and-quarantine, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published External query-routing source for predicted query difficulty, small/large model routing, dynamic quality-cost tradeoffs, and quality-preserving large-model-call reduction.
ext_routellm_2024 RouteLLM: Learning to Route LLMs with Preference Data external_literature router_learning Unassigned in current structure routing-heads-and-specialist-cores, resource-economics-and-token-budgets, the-efficient-asi-hypothesis, policy-optimization-and-learning-from-feedback source source note available; raw/cache text not published External learned-router source for routing between stronger and weaker LLMs using preference data, cost-quality tradeoffs, and transfer to changed model pairs.
ext_deep_compression_2015 Deep Compression: Compressing Deep Neural Networks with Pruning, Trained Quantization and Huffman Coding external_literature compression_representation Unassigned in current structure compact-generative-systems-and-residual-honesty, rankfold-neuralfold-and-artifact-compression, resource-economics-and-token-budgets source source note available; raw/cache text not published External compression source for pruning, trained quantization, coding, memory-footprint reduction, and speed/energy claims that require reproduction before local evidence use.
ext_lora_2021 LoRA: Low-Rank Adaptation of Large Language Models external_literature compression_representation Unassigned in current structure rankfold-neuralfold-and-artifact-compression, compact-generative-systems-and-residual-honesty, resource-economics-and-token-budgets, policy-optimization-and-learning-from-feedback, coilra-multicoil-rope-and-cyclic-mixers source source note available; raw/cache text not published External low-rank adaptation source for parameter-efficient updates, rank-decomposition adapters, memory reduction, and adaptation-boundary vocabulary.
ext_knowledge_distillation_2015 Distilling the Knowledge in a Neural Network external_literature compression_representation Unassigned in current structure compact-generative-systems-and-residual-honesty, rankfold-neuralfold-and-artifact-compression, resource-economics-and-token-budgets source source note available; raw/cache text not published External compression source for teacher/student distillation, soft-target transfer, ensemble compression, and knowledge-transfer claims requiring local reproduction before evidence use.
ext_gptq_2022 GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers external_literature compression_quantization Unassigned in current structure rankfold-neuralfold-and-artifact-compression, compact-generative-systems-and-residual-honesty, resource-economics-and-token-budgets, fast-generation-architectures source source note available; raw/cache text not published External quantization source for post-training compression of large generative transformers, one-shot weight quantization, memory reduction, and accuracy/speed tradeoff boundaries.
ext_qlora_2023 QLoRA: Efficient Finetuning of Quantized LLMs external_literature compression_quantized_adaptation prototype-roadmap (Prototype Roadmap) rankfold-neuralfold-and-artifact-compression, resource-economics-and-token-budgets, policy-optimization-and-learning-from-feedback, prototype-roadmap source source note available; raw/cache text not published External quantized-adaptation source for finetuning quantized LLMs with low-rank adapters, memory-efficient training, and benchmark claims requiring reproduction before local evidence use.
ext_dreamcoder_2020 DreamCoder: Growing generalizable, interpretable knowledge with wake-sleep Bayesian program learning external_literature program_synthesis_representation cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) cognitive-compilation-and-semantic-ir, procedural-memory-and-cognitive-loop-closure, compact-generative-systems-and-residual-honesty, rankfold-neuralfold-and-artifact-compression, mathematical-and-search-substrates, replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published External program-synthesis source for wake-sleep library learning, reusable abstractions, interpretable learned programs, and compression-through-abstraction vocabulary.
ext_llvm_langref_docs LLVM Language Reference Manual external_literature compiler_ir cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR) cognitive-compilation-and-semantic-ir, executable-specifications-and-lean-proof-envelope source source note available; raw/cache text not published Official LLVM Language Reference comparator for SSA-based intermediate representation, equivalent in-memory/bitcode/human-readable forms, well-formedness, verifier passes, and optimization/analysis vocabulary.
ext_mlir_2020 MLIR: A Compiler Infrastructure for the End of Moore’s Law external_literature multi_level_compiler_ir cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR) cognitive-compilation-and-semantic-ir, resource-economics-and-token-budgets, mathematical-and-search-substrates source source note available; raw/cache text not published External multi-level compiler-IR comparator for reusable and extensible compiler infrastructure, dialects, progressive lowering, verifiers, modular passes, and heterogeneous target support.
ext_translation_validation_1998 Translation Validation external_literature translation_validation cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR) cognitive-compilation-and-semantic-ir, executable-specifications-and-lean-proof-envelope source source note available; raw/cache text not published External translation-validation comparator for checking each compiler/code-generator run after translation, using a common semantic framework, refinement relation, and simulation-based proof method.
ext_toolformer_2023 Toolformer: Language Models Can Teach Themselves to Use Tools external_literature learned_tool_use procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure) procedural-memory-and-cognitive-loop-closure, runtime-adapters-tool-permissions-and-human-approval, policy-optimization-and-learning-from-feedback source source note available; raw/cache text not published External learned-tool-use source for self-supervised API-call insertion, tool selection, argument construction, and result incorporation without claiming ASI Stack tool-use reproduction.
ext_voyager_2023 Voyager: An Open-Ended Embodied Agent with Large Language Models external_literature lifelong_skill_learning open-ended-improvement-engines (Open-Ended Improvement Engines); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure) procedural-memory-and-cognitive-loop-closure, routing-heads-and-specialist-cores, benchmark-ratchets-and-anti-goodhart-evidence, open-ended-improvement-engines source source note available; raw/cache text not published External lifelong-agent source for automatic curriculum, executable-code skill libraries, iterative environment-feedback prompting, self-verification, and skill-library transfer in Minecraft.
ext_information_bottleneck_2000 The information bottleneck method external_literature representation_compression learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) compact-generative-systems-and-residual-honesty, resource-economics-and-token-budgets source source note available; raw/cache text not published External representation-compression source for relevance-preserving compression, bottleneck variables, mutual-information tradeoffs, and compression/utility separation.
ext_mdl_tutorial_2004 A tutorial introduction to the minimum description length principle external_literature description_length_residuals learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) compact-generative-systems-and-residual-honesty, rankfold-neuralfold-and-artifact-compression, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published External description-length source for model/data tradeoffs, compression as inductive discipline, and residual/error-accounting vocabulary.
ext_weakness_generalization_2023 The Optimal Choice of Hypothesis Is the Weakest, Not the Shortest external_literature learning_theory_inductive_bias learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) learning-theory-generalization-and-scaling-science source source note available; raw/cache text not published Bennett’s finite enactive-cognition formalism separates extension-based hypothesis weakness from description length. Under a uniform distribution over its task space, the paper argues that maximizing weakness is necessary and sufficient for maximizing generalization probability and gives a counterexample to MDL as a universal proxy. Its theorem assumptions and toy 8-bit arithmetic experiments do not establish a general result for neural networks or real task distributions.
ext_codebleu_2020 CodeBLEU: a Method for Automatic Evaluation of Code Synthesis external_literature artifact_utility_metrics prototype-roadmap (Prototype Roadmap) benchmark-ratchets-and-anti-goodhart-evidence, compact-generative-systems-and-residual-honesty, artifact-steward-agents-and-living-project-governance, prototype-roadmap source source note available; raw/cache text not published External code-synthesis evaluation source for combining lexical, syntax, data-flow, and semantic matching into artifact-quality metrics that still require task-specific validation.
ext_mmlu_2020 Measuring Massive Multitask Language Understanding external_literature benchmark_science prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) benchmark-ratchets-and-anti-goodhart-evidence, evidence-states-and-claim-discipline, open-research-agenda-and-bibliography-plan, prototype-roadmap source source note available; raw/cache text not published External benchmark source for broad multitask evaluation, task-coverage limits, lopsided performance, uncertainty about wrong answers, and benchmark-saturation pressure.
ext_bigbench_2022 Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models external_literature benchmark_science Unassigned in current structure benchmark-ratchets-and-anti-goodhart-evidence, the-efficient-asi-hypothesis, readiness-gates-residual-escrow-and-quarantine, policy-optimization-and-learning-from-feedback source source note available; raw/cache text not published External benchmark source for BIG-bench, broad task coverage, scale effects, calibration, breakthrough behavior, human-rater baselines, and social-bias tradeoffs.
ext_helm_2022 Holistic Evaluation of Language Models external_literature benchmark_science living-book-methodology (Living Book Methodology) benchmark-ratchets-and-anti-goodhart-evidence, evidence-states-and-claim-discipline, readiness-gates-residual-escrow-and-quarantine, living-book-methodology source source note available; raw/cache text not published External benchmark-science source for multi-scenario, multi-metric evaluation, missing-coverage disclosure, raw-prompt transparency, and living benchmark practice.
ext_gpqa_2023 GPQA: A Graduate-Level Google-Proof Q&A Benchmark external_literature benchmark_science verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) benchmark-ratchets-and-anti-goodhart-evidence, verification-bandwidth-and-context-adequacy, evidence-states-and-claim-discipline, open-research-agenda-and-bibliography-plan source source note available; raw/cache text not published External benchmark source for expert-written hard questions, Google-proof validation, scalable oversight pressure, and the gap between skilled non-expert validation and expert competence.
ext_swe_bench_2023 SWE-bench: Can Language Models Resolve Real-World GitHub Issues? external_literature benchmark_science prototype-roadmap (Prototype Roadmap) benchmark-ratchets-and-anti-goodhart-evidence, artifact-graphs-audit-logs-and-replay, labor-os-and-typed-jobs, prototype-roadmap source source note available; raw/cache text not published External benchmark source for real-world software-engineering issue resolution, repository-scale context, executable environments, patch evaluation, and capability boundaries.
ext_swe_rebench_v2_2026 SWE-rebench V2: Language-Agnostic SWE Task Collection at Scale external_literature natural_software_task_construction_and_evaluation artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap) benchmark-ratchets-and-anti-goodhart-evidence, artifact-graphs-audit-logs-and-replay, integrated-reference-architecture, prototype-roadmap, readiness-gates-residual-escrow-and-quarantine source source note available; raw/cache text not published Primary 2026 natural-task substrate for multilingual repository changes, interactive setup, containerized full-suite execution, separated solution/test patches, task diagnostics, and explicit environment pathologies. It does not establish local task validity, gold execution, model competence, governance benefit, safety, transfer, or SOTA.
ext_livebench_2024 LiveBench: A Challenging, Contamination-Limited LLM Benchmark external_literature benchmark_science living-book-methodology (Living Book Methodology); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) benchmark-ratchets-and-anti-goodhart-evidence, evidence-states-and-claim-discipline, living-book-methodology, open-research-agenda-and-bibliography-plan source source note available; raw/cache text not published External benchmark source for contamination-limited evaluation, frequently updated questions, objective ground-truth scoring, and monthly benchmark evolution.
ext_dynabench_2021 Dynabench: Rethinking Benchmarking in NLP external_literature dynamic_benchmarking Unassigned in current structure benchmark-ratchets-and-anti-goodhart-evidence, readiness-gates-residual-escrow-and-quarantine, policy-optimization-and-learning-from-feedback, artifact-steward-agents-and-living-project-governance source source note available; raw/cache text not published External dynamic-benchmarking source for human-and-model-in-the-loop data collection, adversarial benchmark evolution, and stale static benchmark pressure.
ext_checklist_2020 Beyond Accuracy: Behavioral Testing of NLP models with CheckList external_literature behavioral_evaluation verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); prototype-roadmap (Prototype Roadmap) benchmark-ratchets-and-anti-goodhart-evidence, verification-bandwidth-and-context-adequacy, claim-ledgers-and-belief-revision, prototype-roadmap source source note available; raw/cache text not published External behavioral-testing source for capability matrices, minimum functionality tests, invariance tests, directional expectation tests, and failure-discovery beyond aggregate accuracy.
ext_benchmark_contamination_2023 Investigating Data Contamination in Modern Benchmarks for Large Language Models external_literature benchmark_contamination living-book-methodology (Living Book Methodology) benchmark-ratchets-and-anti-goodhart-evidence, evidence-states-and-claim-discipline, readiness-gates-residual-escrow-and-quarantine, living-book-methodology source source note available; raw/cache text not published External benchmark-contamination source for detecting training/test overlap pressure, benchmark-leakage risk, and score interpretation limits in modern LLM evaluations.
ext_goodhart_variants_2018 Categorizing Variants of Goodhart’s Law external_literature goodhart_failure_taxonomy failure-modes-of-ungoverned-intelligence (Failure Modes of Ungoverned Intelligence) failure-modes-of-ungoverned-intelligence, benchmark-ratchets-and-anti-goodhart-evidence, policy-optimization-and-learning-from-feedback, artifact-steward-agents-and-living-project-governance, evidence-states-and-claim-discipline source source note available; raw/cache text not published External Goodhart-taxonomy source for regressive, extremal, causal, and adversarial metric failures that benchmark ratchets and policy updates must treat as distinct risks.
ext_speculative_decoding_2022 Fast Inference from Transformers via Speculative Decoding external_literature fast_generation fast-generation-architectures (Fast Generation Architectures) fast-generation, the-efficient-asi-hypothesis source source note available; raw/cache text not published Primary external paper for speculative decoding: a draft model proposes multiple tokens and a target model verifies them, giving an exact-distribution acceleration path under its assumptions.
ext_multi_token_prediction_2024 Better & Faster Large Language Models via Multi-token Prediction external_literature fast_generation fast-generation-architectures (Fast Generation Architectures) fast-generation, the-efficient-asi-hypothesis source source note available; raw/cache text not published Primary external paper for multi-token prediction as an auxiliary training objective and inference-time multi-token proposal mechanism.
ext_medusa_2024 Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads external_literature fast_generation fast-generation-architectures (Fast Generation Architectures) fast-generation, the-efficient-asi-hypothesis source source note available; raw/cache text not published Primary external paper for adding multiple decoding heads to an LLM and verifying tree-structured candidate continuations in parallel.
ext_eagle_2024 EAGLE: Speculative Sampling Requires Rethinking Feature Uncertainty external_literature fast_generation fast-generation-architectures (Fast Generation Architectures) fast-generation, the-efficient-asi-hypothesis source source note available; raw/cache text not published Primary external paper for feature-level speculative drafting and target-model verification as an acceleration mechanism.
ext_lookahead_decoding_2024 Break the Sequential Dependency of LLM Inference Using Lookahead Decoding external_literature fast_generation fast-generation-architectures (Fast Generation Architectures) fast-generation source source note available; raw/cache text not published Primary external paper for lookahead decoding: a parallel exact decoding algorithm that reduces serial decoding steps without an auxiliary draft model.
ext_layerskip_2024 LayerSkip: Enabling Early Exit Inference and Self-Speculative Decoding external_literature fast_generation fast-generation-architectures (Fast Generation Architectures) fast-generation source source note available; raw/cache text not published Primary external paper for early-exit inference and self-speculative decoding where early layers draft and later layers verify.
ext_pagedattention_vllm_2023 Efficient Memory Management for Large Language Model Serving with PagedAttention external_literature fast_generation virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation, resource-economics source source note available; raw/cache text not published Primary external paper for vLLM/PagedAttention, which treats KV-cache memory management and serving throughput as a distinct acceleration axis.
ext_transformer_xl_2019 Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context external_literature sequence_memory_recurrence Unassigned in current structure coil-attention-cyclic-memory-and-recurrence-contracts source source note available; raw/cache text not published External recurrent Transformer comparator for segment-level recurrence, relative positional encoding, and long-dependency language modeling; useful as a baseline family for cyclic-memory contracts without implying local reproduction.
ext_compressive_transformer_2019 Compressive Transformers for Long-Range Sequence Modelling external_literature sequence_memory_recurrence Unassigned in current structure coil-attention-cyclic-memory-and-recurrence-contracts source source note available; raw/cache text not published External long-range memory comparator for compressed past memories, memory mechanisms, and long-range sequence benchmarks; useful for positioning cyclic memory against compression-memory baselines.
ext_roformer_rope_2021 RoFormer: Enhanced Transformer with Rotary Position Embedding external_literature position_encoding Unassigned in current structure coilra-multicoil-rope-and-cyclic-mixers source source note available; raw/cache text not published External RoPE comparator for rotary position embedding, relative-position behavior inside self-attention, and position-encoding baselines for cyclic phase or RoPE-style substrates.
ext_retnet_2023 Retentive Network: A Successor to Transformer for Large Language Models external_literature sequence_memory_recurrence replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) coil-attention-cyclic-memory-and-recurrence-contracts, coilra-multicoil-rope-and-cyclic-mixers, replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published External retention/recurrent-sequence comparator for the relationship between recurrence and attention, recurrent/chunkwise computation, and inference-efficiency tradeoffs.
ext_mamba_2023 Mamba: Linear-Time Sequence Modeling with Selective State Spaces external_literature sequence_substrates replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); fast-generation-architectures (Fast Generation Architectures); mathematical-and-search-substrates (Mathematical and Search Substrates) fast-generation, math-substrates, coilra-multicoil-rope-and-cyclic-mixers, replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary external paper for selective state-space sequence models as a different long-sequence substrate and inference-efficiency axis from decoding tricks.
ext_llada_2025 Large Language Diffusion Models external_literature diffusion_language_models replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); fast-generation-architectures (Fast Generation Architectures) fast-generation-architectures, replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary external paper for LLaDA, a large masked-diffusion language model trained with pretraining and supervised fine-tuning rather than left-to-right autoregression.
ext_scaling_dllms_2026 Scaling Beyond Masked Diffusion Language Models external_literature diffusion_language_models replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); fast-generation-architectures (Fast Generation Architectures) fast-generation-architectures, replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary external paper for comparing diffusion language-model families by speed-quality tradeoffs rather than perplexity alone.
ext_trpo_2015 Trust Region Policy Optimization external_literature policy_optimization policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) governed-deliberation-and-test-time-scaling source source note available; raw/cache text not published Primary external source for trust-region policy-gradient updates and bounded update-size discipline.
ext_ppo_2017 Proximal Policy Optimization Algorithms external_literature policy_optimization policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) governed-deliberation-and-test-time-scaling source source note available; raw/cache text not published Primary external source for PPO-style online policy-gradient updates and proximal surrogate objectives.
ext_remax_2023 ReMax: A Simple, Effective, and Efficient Reinforcement Learning Method for Aligning Large Language Models external_literature policy_optimization policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) source source note available; raw/cache text not published Primary external source for simpler RLHF-oriented policy-gradient updates relative to PPO-style machinery.
ext_goal_oriented_requirements_engineering_2001 Goal-Oriented Requirements Engineering: A Guided Tour external_literature requirements_engineering human-intent-as-a-formal-input (Human Intent as a Formal Input) human-intent-as-a-formal-input source source note available; raw/cache text not published External requirements-engineering comparator for turning stakeholder goals, constraints, refinements, and responsibilities into explicit requirements before system design or execution.
ext_cooperative_inverse_rl_2016 Cooperative Inverse Reinforcement Learning external_literature human_intent_alignment human-intent-as-a-formal-input (Human Intent as a Formal Input); governed-objective-formation-value-learning-and-goal-integrity (Governed Objective Formation, Value Learning, and Goal Integrity) human-intent-as-a-formal-input source source note available; raw/cache text not published External cooperative AI comparator for formalizing value alignment as uncertainty about the human reward function in a cooperative partial-information setting.
ext_deep_rl_human_preferences_2017 Deep Reinforcement Learning from Human Preferences external_literature human_feedback_learning human-intent-as-a-formal-input (Human Intent as a Formal Input) human-intent-as-a-formal-input, policy-optimization-and-learning-from-feedback source source note available; raw/cache text not published External human-feedback comparator for communicating complex goals through preference comparisons over behavior segments; useful for separating preference signals from explicit intent contracts.
ext_dpo_2023 Direct Preference Optimization: Your Language Model is Secretly a Reward Model external_literature policy_optimization policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) source source note available; raw/cache text not published Primary external source for DPO-style offline preference optimization without a separate online RL loop.
ext_ipo_preference_2023 A General Theoretical Paradigm to Understand Learning from Human Preferences external_literature policy_optimization policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) source source note available; raw/cache text not published External source for theoretical comparison of preference-learning objectives, including IPO/DPO-style framing.
ext_orpo_2024 ORPO: Monolithic Preference Optimization without Reference Model external_literature policy_optimization policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) source source note available; raw/cache text not published External source for reference-model-free monolithic preference optimization.
ext_kto_2024 KTO: Model Alignment as Prospect Theoretic Optimization external_literature policy_optimization policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) source source note available; raw/cache text not published External source for prospect-theoretic model alignment and human-aware loss framing.
ext_simpo_2024 SimPO: Simple Preference Optimization with a Reference-Free Reward external_literature policy_optimization policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) source source note available; raw/cache text not published External source for simple reference-free preference optimization using sequence-level reward framing.
ext_reinforce_style_rlhf_2024 Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs external_literature policy_optimization policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) source source note available; raw/cache text not published External source for revisiting simpler REINFORCE-style optimization as an RLHF baseline.
ext_deepseek_r1_2025 DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning external_literature policy_optimization governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) source source note available; raw/cache text not published External source for reinforcement-learning pressure on reasoning behavior in large language models.
ext_dapo_2025 DAPO: An Open-Source LLM Reinforcement Learning System at Scale external_literature policy_optimization policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) source source note available; raw/cache text not published External source for open-source large-scale LLM RL system details and DAPO-style update design.
ext_gspo_2025 Group Sequence Policy Optimization external_literature policy_optimization policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) source source note available; raw/cache text not published External source for sequence-level group policy optimization in LLM reinforcement learning.
ext_s_grpo_2025 S-GRPO: Early Exit via Reinforcement Learning in Reasoning Models external_literature policy_optimization governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) source source note available; raw/cache text not published External source for early-exit reinforcement learning and overthinking control in reasoning models.
ext_longrlvr_2026 LongRLVR: Long-Context Reinforcement Learning Requires Verifiable Context Rewards external_literature policy_optimization policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) source source note available; raw/cache text not published External source for long-context RLVR and verifiable context-grounding rewards.
ext_rlhf_limitations_2023 Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback external_literature policy_optimization policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) source source note available; raw/cache text not published External survey source for RLHF limitations, reward hacking, evaluator limits, and complementary safeguards.
ext_tailscale_docs_2025 What is Tailscale? external_literature personal_compute_hives personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) personal-compute-hives-and-federated-edge-intelligence source source note available; raw/cache text not published Official Tailscale documentation for zero-trust identity networking, tailnets, encrypted point-to-point connections, and cross-network device connectivity.
ext_kubernetes_overview_docs Kubernetes Documentation: Overview external_literature personal_compute_hives personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) personal-compute-hives-and-federated-edge-intelligence source source note available; raw/cache text not published Official Kubernetes overview for containerized workload management, declarative configuration, automation, service discovery, storage orchestration, rollouts, bin packing, and self-healing.
ext_k3s_docs_2026 K3s: Lightweight Kubernetes external_literature personal_compute_hives personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) personal-compute-hives-and-federated-edge-intelligence source source note available; raw/cache text not published Official K3s documentation for lightweight Kubernetes deployment in edge, homelab, IoT, CI, single-board-computer, air-gapped, and embedded settings.
ext_nomad_docs Nomad Documentation external_literature personal_compute_hives personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) personal-compute-hives-and-federated-edge-intelligence source source note available; raw/cache text not published Official Nomad documentation for scheduling and orchestrating containers, non-containerized applications, and batch jobs across on-prem and cloud environments.
ext_temporal_docs Temporal Documentation: What is Temporal? external_literature durable_execution labor-os-and-typed-jobs (Labor OS and Typed Jobs) intent-to-execution-contracts, labor-os-and-typed-jobs source source note available; raw/cache text not published Official Temporal documentation comparator for durable workflow execution, workflow event histories, worker processes, failure recovery, and long-running application-code orchestration.
ext_airflow_dag_docs Apache Airflow Documentation: Dags external_literature workflow_orchestration labor-os-and-typed-jobs (Labor OS and Typed Jobs) intent-to-execution-contracts, labor-os-and-typed-jobs source source note available; raw/cache text not published Official Apache Airflow documentation comparator for DAG-based workflow scheduling, tasks, dependencies, callbacks, retries, and operational workflow metadata.
ext_bpmn_2_0_2_spec Business Process Model and Notation Specification Version 2.0.2 external_literature business_process_modeling labor-os-and-typed-jobs (Labor OS and Typed Jobs) intent-to-execution-contracts, labor-os-and-typed-jobs source source note available; raw/cache text not published OMG BPMN 2.0.2 formal specification comparator for stakeholder-readable business-process diagrams, implementation-independent flow notation, and translation into software process components.
ext_kubernetes_jobs_docs Kubernetes Documentation: Jobs external_literature batch_job_lifecycle labor-os-and-typed-jobs (Labor OS and Typed Jobs) labor-os-and-typed-jobs source source note available; raw/cache text not published Official Kubernetes Jobs documentation comparator for batch job lifecycle, completions, backoff limits, active deadlines, terminal Complete/Failed conditions, and cleanup of finished jobs.
ext_ray_core_docs_2026 What’s Ray Core? external_literature personal_compute_hives personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) personal-compute-hives-and-federated-edge-intelligence source source note available; raw/cache text not published Official Ray Core documentation for distributed task, actor, and object primitives used to build and scale Python applications.
ext_boinc_home_2026 BOINC external_literature personal_compute_hives personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) personal-compute-hives-and-federated-edge-intelligence source source note available; raw/cache text not published Official BOINC site for volunteer computing where user computers download scientific computing jobs and run them in the background.
ext_syncthing_home Syncthing external_literature personal_compute_hives personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) personal-compute-hives-and-federated-edge-intelligence source source note available; raw/cache text not published Official Syncthing site for continuous file synchronization across computers, authenticated devices, encrypted transport, and user-controlled storage location.
ext_ipfs_docs IPFS Documentation and Project Site external_literature personal_compute_hives personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence) personal-compute-hives-and-federated-edge-intelligence source source note available; raw/cache text not published Official IPFS project documentation and site for peer-to-peer content addressing, content identifiers, provider discovery, and decentralized retrieval vocabulary.
ext_akash_docs_2026 Akash Network Documentation external_literature personal_compute_hives personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) personal-compute-hives-and-federated-edge-intelligence, artifact-steward-agents-and-living-project-governance source source note available; raw/cache text not published Official Akash documentation for decentralized cloud deployment, provider resources, leases, GPUs, SDKs, node operations, and provider operations.
ext_golem_docs_2025 Golem Developer Resources external_literature personal_compute_hives personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) personal-compute-hives-and-federated-edge-intelligence, artifact-steward-agents-and-living-project-governance source source note available; raw/cache text not published Official Golem developer resources for decentralized computations, task execution, provider selection, result handling, and resource sharing.
ext_github_webhooks_docs Webhook events and payloads external_literature artifact_steward_agents artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) artifact-steward-agents-and-living-project-governance source source note available; raw/cache text not published Official GitHub documentation for repository and organization webhook events, event payloads, delivery headers, event-specific permissions, and payload limits.
ext_github_self_hosted_runners_docs Self-hosted runners external_literature artifact_steward_agents personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) artifact-steward-agents-and-living-project-governance, personal-compute-hives-and-federated-edge-intelligence source source note available; raw/cache text not published Official GitHub Actions documentation for self-hosted runners as user-managed systems that execute workflow jobs on physical, virtual, containerized, on-prem, or cloud machines.
ext_openzeppelin_governor_docs OpenZeppelin Contracts: Governance external_literature artifact_steward_agents artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) artifact-steward-agents-and-living-project-governance source source note available; raw/cache text not published Official OpenZeppelin governance documentation for modular Governor contracts, voting power, quorum, timelocks, proposal settings, and guardian-style extensions.
ext_open_collective_docs Open Collective Documentation external_literature artifact_steward_agents artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) artifact-steward-agents-and-living-project-governance source source note available; raw/cache text not published Official Open Collective documentation for transparent community money management, fiscal hosts, contribution intake, expenses, accounting, and legal-entity delegation through fiscal hosting.
ext_github_sponsors_docs About GitHub Sponsors for open source contributors external_literature artifact_steward_agents artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) artifact-steward-agents-and-living-project-governance source source note available; raw/cache text not published Official GitHub Sponsors documentation for contributor and organization sponsorship eligibility, open-source contribution categories, sponsor profiles, and GitHub-native funding surfaces.
ext_agentic_workflow_injection_2026 Demystifying and Detecting Agentic Workflow Injection Vulnerabilities in GitHub Actions external_literature artifact_steward_agents artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) artifact-steward-agents-and-living-project-governance source source note available; raw/cache text not published External security paper for agentic workflow injection in GitHub Actions when untrusted repository event context reaches LLM agents and downstream workflow logic.
ext_dao_delegation_fairness_2025 Fairness in Token Delegation: Mitigating Voting Power Concentration in DAOs external_literature artifact_steward_agents artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) artifact-steward-agents-and-living-project-governance source source note available; raw/cache text not published External DAO governance paper for voter apathy, voting-power concentration, delegation misalignment, and delegate-ranking bias.
ext_model_cards_2019 Model Cards for Model Reporting external_literature model_reporting project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference) evidence-states-and-claim-discipline, project-theseus-as-report-first-implementation-reference source source note available; raw/cache text not published External reporting comparator for structured model documentation, intended-use boundaries, evaluation disclosures, ethical considerations, and model-report artifacts.
ext_datasheets_datasets_2021 Datasheets for Datasets external_literature dataset_documentation project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference) evidence-states-and-claim-discipline, project-theseus-as-report-first-implementation-reference source source note available; raw/cache text not published External documentation comparator for dataset motivation, composition, collection, preprocessing, uses, distribution, maintenance, and accountability questions.
ext_factsheets_ai_services_2019 FactSheets: Increasing Trust in AI Services through Supplier’s Declarations of Conformity external_literature ai_service_fact_sheets project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference) project-theseus-as-report-first-implementation-reference source source note available; raw/cache text not published External AI-service documentation comparator for supplier declarations, service properties, trust-relevant facts, and standardized reporting boundaries.
ext_ml_reproducibility_program_2021 Improving Reproducibility in Machine Learning Research external_literature ml_reproducibility_reporting project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference) evidence-states-and-claim-discipline, project-theseus-as-report-first-implementation-reference source source note available; raw/cache text not published External reproducibility-program comparator for checklists, code submission, reproducibility reports, and community review mechanisms in machine-learning research.
ext_transformer_circuits_2021 A Mathematical Framework for Transformer Circuits external_literature mechanistic_interpretability white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) evidence-states-and-claim-discipline, white-box-evidence-interpretability-and-activation-governance source source note available; raw/cache text not published External mechanistic-interpretability comparator for treating internal circuit analyses as scoped white-box evidence that still needs model, layer, behavior, causal, and limitation boundaries.
ext_monosemanticity_2023 Towards Monosemanticity: Decomposing Language Models With Dictionary Learning external_literature mechanistic_interpretability white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) evidence-states-and-claim-discipline, white-box-evidence-interpretability-and-activation-governance source source note available; raw/cache text not published External mechanistic-interpretability comparator for sparse-autoencoder feature decomposition and the boundary between discovered features, feature-level evidence, and broader model-behavior claims.
ext_literate_programming_1984 Literate Programming external_literature literate_programming living-book-methodology (Living Book Methodology) living-book-methodology source source note available; raw/cache text not published External literate-programming source for arranging programs and explanation around human comprehension, woven documentation, and tangled executable artifacts.
ext_jupyter_book_docs Jupyter Book Documentation external_literature executable_books living-book-methodology (Living Book Methodology) living-book-methodology source source note available; raw/cache text not published Official Jupyter Book documentation comparator for authoring books from notebooks or Markdown, executing code, cross-referencing content, and publishing computational books to the web.
ext_quarto_books_docs Quarto Books Documentation external_literature technical_publishing living-book-methodology (Living Book Methodology) living-book-methodology source source note available; raw/cache text not published Official Quarto Books documentation comparator for multi-chapter manuscripts, HTML/PDF/Word/EPUB outputs, search, cross references, and book-style website publishing.
ext_argo_rollouts_docs Argo Rollouts Documentation: Kubernetes Progressive Delivery Controller external_literature progressive_delivery_rollback capability-replacement-and-rollback (Capability Replacement and Rollback) capability-replacement-and-rollback source source note available; raw/cache text not published External progressive-delivery comparator for blue-green rollout, canary rollout, metric analysis, automated promotion, and automated rollback vocabulary.
ext_feature_toggles_fowler Feature Toggles (aka Feature Flags) external_literature feature_flag_release_control capability-replacement-and-rollback (Capability Replacement and Rollback) capability-replacement-and-rollback source source note available; raw/cache text not published External feature-flag comparator for controlled exposure, canary releasing, release toggles, experiment toggles, ops toggles, permissioning toggles, and validation complexity.
ext_google_cloud_mlops_cd MLOps: Continuous Delivery and Automation Pipelines in Machine Learning external_literature mlops_continuous_delivery capability-replacement-and-rollback (Capability Replacement and Rollback) capability-replacement-and-rollback source source note available; raw/cache text not published External MLOps comparator for CI/CD/CT, data/model validation, model deployment, monitoring, rollback triggers, and model-regression concerns.
ext_kubernetes_deployments_docs Kubernetes Documentation: Deployments external_literature deployment_rollout_rollback capability-replacement-and-rollback (Capability Replacement and Rollback) capability-replacement-and-rollback source source note available; raw/cache text not published External deployment-controller comparator for rollout status, rollout history, revision records, and rollback to a prior stable Deployment revision.
ext_drexler_cais_2019 Reframing Superintelligence: Comprehensive AI Services as General Intelligence external_literature ai_services_r_and_d_automation asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); integrated-reference-architecture (Integrated Reference Architecture) asi-is-a-stack-not-a-model, constitutional-alignment-substrate, recursive-self-improvement-boundaries, integrated-reference-architecture source source note available; raw/cache text not published Primary CAIS technical-report comparator for service-centered general intelligence, R&D automation, structured AI development, and the distinction between component affordances and a complete governance architecture.
ext_mrkl_systems_2022 MRKL Systems: A Modular, Neuro-Symbolic Architecture That Combines Large Language Models, External Knowledge Sources and Discrete Reasoning external_literature neuro_symbolic_modular_architecture Unassigned in current structure asi-is-a-stack-not-a-model source source note available; raw/cache text not published External modular-neuro-symbolic architecture comparator for combining language models with expert modules, external knowledge sources, and routing rather than treating the model as the whole system.
ext_llm_agents_survey_2023 A Survey on Large Language Model based Autonomous Agents external_literature llm_agent_architecture Unassigned in current structure asi-is-a-stack-not-a-model source source note available; raw/cache text not published External LLM-agent architecture comparator for agent profiles, memory, planning, and action modules around a language model, useful for positioning the stack frame against agent-system decompositions.
ext_standard_model_mind_2017 A Standard Model of the Mind: Toward a Common Computational Framework Across Artificial Intelligence, Cognitive Science, Neuroscience, and Robotics external_literature cognitive_architecture Unassigned in current structure asi-is-a-stack-not-a-model source source note available; raw/cache text not published External cognitive-architecture comparator for treating intelligent behavior as an integrated architecture spanning memory, learning, perception/action, procedural control, and deliberation.
ext_subsumption_architecture_1986 A Robust Layered Control System for a Mobile Robot external_literature layered_robot_control_architecture Unassigned in current structure asi-is-a-stack-not-a-model source source note available; raw/cache text not published External layered-control architecture comparator for decomposing robot behavior into interacting layers rather than centralizing behavior in one monolithic controller.
ext_humans_automation_1997 Humans and Automation: Use, Misuse, Disuse, Abuse external_literature human_factors_automation human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) runtime-adapters-tool-permissions-and-human-approval, human-intent-as-a-formal-input, evidence-states-and-claim-discipline, human-factors-and-meaningful-control-in-oversight source source note available; raw/cache text not published External human-factors comparator for automation use, misuse, disuse, abuse, overreliance, monitoring failure, workload, trust, risk, false alarms, and operator-role design.
ext_ironies_automation_1983 Ironies of Automation external_literature human_factors_automation human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) runtime-adapters-tool-permissions-and-human-approval, evidence-states-and-claim-discipline, human-factors-and-meaningful-control-in-oversight source source note available; raw/cache text not published External automation comparator for the argument that automation can expand rather than eliminate human-operator problems and can leave humans with difficult abnormal-condition duties.
ext_levels_automation_2000 A Model for Types and Levels of Human Interaction with Automation external_literature human_factors_automation human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) runtime-adapters-tool-permissions-and-human-approval, human-intent-as-a-formal-input, human-factors-and-meaningful-control-in-oversight source source note available; raw/cache text not published External human-automation comparator for separating automation by information acquisition, analysis, decision/action selection, and action implementation rather than treating human approval as a single undifferentiated gate.
ext_complacency_bias_automation_2010 Complacency and Bias in Human Use of Automation: An Attentional Integration external_literature human_factors_automation human-factors-and-meaningful-control-in-oversight (Human Factors and Meaningful Control in Oversight); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) runtime-adapters-tool-permissions-and-human-approval, human-intent-as-a-formal-input, evidence-states-and-claim-discipline, human-factors-and-meaningful-control-in-oversight source source note available; raw/cache text not published External human-factors comparator for automation complacency, omission and commission errors, automation bias, workload, attention, and imperfect decision aids.
ext_bourtoule_machine_unlearning_2021 Machine Unlearning external_literature machine_unlearning_data_governance context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) data-engines-continual-learning-and-unlearning, policy-optimization-and-learning-from-feedback, procedural-memory-and-cognitive-loop-closure, benchmark-ratchets-and-anti-goodhart-evidence, adjudicated-persistence-and-the-adaptive-commit-boundary, open-research-agenda-and-bibliography-plan source source note available; raw/cache text not published Primary machine-unlearning comparator for deletion-path architecture, checkpointed training, bounded retraining, accuracy-cost trade-offs, and the boundary between deletion requests and verified removal.
ext_shumailov_model_collapse_2023 The Curse of Recursion: Training on Generated Data Makes Models Forget external_literature synthetic_data_feedback_model_collapse data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) data-engines-continual-learning-and-unlearning, policy-optimization-and-learning-from-feedback, procedural-memory-and-cognitive-loop-closure, benchmark-ratchets-and-anti-goodhart-evidence, open-research-agenda-and-bibliography-plan source source note available; raw/cache text not published Primary preprint comparator for generated-data feedback, provenance, distribution-tail loss, and model-collapse risk under specified recursive-training assumptions; not a universal synthetic-data safety result.
ext_gerstgrasser_data_accumulation_2024 Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data external_literature synthetic_data_retention_policy data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) data-engines-continual-learning-and-unlearning, policy-optimization-and-learning-from-feedback, procedural-memory-and-cognitive-loop-closure, benchmark-ratchets-and-anti-goodhart-evidence, open-research-agenda-and-bibliography-plan source source note available; raw/cache text not published Primary empirical and analytical comparator that distinguishes replacement from accumulation of real and synthetic data; gives a counterweight to blanket model-collapse claims without resolving deletion, privacy, provenance, or poisoning risk.
theseus_synthetic_data_curation Theseus Synthetic Data Curation implementation_reference governed_synthetic_data_admission privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) benchmark-ratchets-and-anti-goodhart-evidence, data-engines-continual-learning-and-unlearning, policy-optimization-and-learning-from-feedback, privacy-data-rights-and-information-flow-governance, procedural-memory-and-cognitive-loop-closure, project-theseus-as-report-first-implementation-reference local project reference (not publicly linked) source note available; raw/cache text not published Pinned Project Theseus implementation-reference record for residual-targeted synthetic-data admission, provenance, leakage checks, ratio caps, governed teacher handling, and source-reported promotion gates; no ASI Stack replay or model-quality import.
ext_alignment_faking_2024 Alignment Faking in Large Language Models external_literature training_time_deception adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) adversarial-evaluation-sandbagging-and-training-time-deception source source note available; raw/cache text not published Primary experimental comparator for context-dependent alignment faking under disclosed training conditions; does not establish that an ASI Stack model, evaluator, or training process is deceptive.
ext_ai_sandbagging_2024 AI Sandbagging: Language Models Can Strategically Underperform on Evaluations external_literature evaluation_integrity adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) adversarial-evaluation-sandbagging-and-training-time-deception source source note available; raw/cache text not published Primary experimental comparator for strategic evaluation underperformance, prompted and password-locked capability hiding, and limits of capability-evaluation trustworthiness; does not show sandbagging in this repository.
ext_emergent_misalignment_reward_hacking_2025 Natural Emergent Misalignment from Reward Hacking in Production RL external_literature training_time_deception inner-alignment-mesa-optimization-and-learned-objective-integrity (Inner Alignment, Mesa-Optimization, and Learned-Objective Integrity); governed-objective-formation-value-learning-and-goal-integrity (Governed Objective Formation, Value Learning, and Goal Integrity); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) adversarial-evaluation-sandbagging-and-training-time-deception source source note available; raw/cache text not published Primary experimental comparator for reward-hacking-induced misaligned generalization in a specified production-RL research setting, including reported mitigation conditions; it is not evidence of local model behavior or a general causal law.
ext_poet_2019 Paired Open-Ended Trailblazer (POET): Endlessly Generating Increasingly Complex and Diverse Learning Environments and Their Solutions external_literature open_ended_environment_solution_generation open-ended-improvement-engines (Open-Ended Improvement Engines) open-ended-improvement-engines source source note available; raw/cache text not published Primary open-ended-learning comparator for paired environment generation, agent optimization, and cross-environment solution transfer in a specified BipedalWalker setting; it does not establish a general improvement engine, evaluator soundness, or ASI Stack result.
ext_funsearch_2024 Mathematical Discoveries from Program Search with Large Language Models external_literature evaluator_bounded_program_search open-ended-improvement-engines (Open-Ended Improvement Engines) open-ended-improvement-engines source source note available; raw/cache text not published Primary program-search comparator for a fixed pretrained LLM, user-provided evaluation function, candidate program archive, and iterative search over a bounded specification; it does not establish open-ended general intelligence, self-modification, evaluator correctness, or an ASI Stack result.
ext_gsn_community_standard_2011 GSN Community Standard Version 1 external_literature structured_assurance_argumentation safety-cases-and-structured-assurance (Safety Cases and Structured Assurance) safety-cases-and-structured-assurance source source note available; raw/cache text not published Primary notation standard comparator for explicit goals, strategies, solutions, context, assumptions, justifications, and relationships in structured assurance arguments; the notation documents asserted support but does not establish claim truth.
ext_evaluations_safety_cases_scheming_2024 Towards Evaluations-Based Safety Cases for AI Scheming external_literature ai_safety_case_methodology safety-cases-and-structured-assurance (Safety Cases and Structured Assurance) safety-cases-and-structured-assurance source source note available; raw/cache text not published Primary safety-case comparator for scoped scheming inability, harm inability, harm control, alignment arguments, empirical evaluation dependencies, and acknowledged unresolved assumptions; it does not establish any ASI Stack safety case or safety result.
ext_aisi_safety_cases_2024 Safety Cases at AISI external_literature ai_safety_case_methodology safety-cases-and-structured-assurance (Safety Cases and Structured Assurance) safety-cases-and-structured-assurance source source note available; raw/cache text not published Official AI Safety Institute methodology comparator for structured safety-case sketches, positive and negative evidence, countercases, open scientific uncertainty, and limits on confidence; it is not evidence that this book has a complete safety case.
ext_rand_model_weight_security_2024 Securing AI Model Weights: Preventing Theft and Misuse of Frontier Models external_literature model_weight_custody model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control) open-weight-release-and-post-release-control, model-weight-custody-and-hardware-roots-of-trust source source note available; raw/cache text not published Primary RAND analysis of frontier-model-weight theft/misuse threat surfaces, security levels, defense-in-depth, access control, physical and organizational controls; it does not establish local protection or safety.
ext_nist_confidential_computing_2026 Hardware-Enabled Security: Confidential Computing of Data in Cloud Workloads external_literature hardware_root_attestation model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust) model-weight-custody-and-hardware-roots-of-trust source source note available; raw/cache text not published NIST initial-public-draft comparator for hardware-enabled confidential computing, memory protection, trust domains, attestation-gated key release, and AI model/data protection; it is draft guidance, not a local TEE result.
ext_nvidia_confidential_model_lifecycle_2026 Workload and Model Lifecycle: Deploying Proprietary Models Securely with NVIDIA Confidential Computing external_literature attestation_gated_model_loading model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust) model-weight-custody-and-hardware-roots-of-trust source source note available; raw/cache text not published Official vendor implementation-reference comparator for encrypted weights outside a confidential pod, policy-sensitive attestation evidence, and key-release decisions; it does not establish a local confidential deployment or attestation result.
ext_provable_model_weight_release_2025 Towards Provable (In)Secure Model Weight Release Schemes external_literature open_weight_release_security model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control) open-weight-release-and-post-release-control, model-weight-custody-and-hardware-roots-of-trust source source note available; raw/cache text not published Primary formal-security comparator for evaluating claimed secure model-weight release schemes and parameter-extraction failure modes; it does not establish an ASI Stack release scheme or release decision.
ext_nist_cscrm_2022 Cybersecurity Supply Chain Risk Management Practices for Systems and Organizations external_literature ai_supply_chain_governance ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance) ai-supply-chain-integrity-and-lifecycle-provenance source source note available; raw/cache text not published Primary NIST C-SCRM standard comparator for lifecycle-wide risk framing, supplier/component inventory, assessment, response, monitoring, and incident communication; it does not establish a local supply-chain program or AI artifact integrity.
ext_slsa_build_track_1_2 SLSA Build Track Basics, version 1.2 external_literature build_provenance ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance) ai-supply-chain-integrity-and-lifecycle-provenance source source note available; raw/cache text not published Official SLSA specification comparator for build provenance, signed hosted builds, verification, and graduated assurance; provenance quality and SLSA level do not establish local artifact correctness, data quality, model safety, or deployment authority.
ext_openssf_model_signing_spec_2025 OpenSSF Model Signing Specification external_literature ai_artifact_signing ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance) ai-supply-chain-integrity-and-lifecycle-provenance source source note available; raw/cache text not published Official OpenSSF AI/ML working-group specification comparator for signed model/dataset bundles, verification, provenance metadata, and explicit limits of model signing; it does not establish local signing, verification, integrity, confidentiality, safety, or release authority.
ext_spdx_ai_profile_3_0_1 SPDX Specification 3.0.1 AI Profile external_literature ai_bill_of_materials ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance) ai-supply-chain-integrity-and-lifecycle-provenance source source note available; raw/cache text not published Official SPDX specification comparator for interoperable AI system/model, dataset, build, supplier, provenance, integrity, relationship, and lifecycle metadata; a conformant BOM is not proof of complete inventory, artifact security, data fitness, model safety, or compliance.
ext_mcp_protocol_2025_06_18 Model Context Protocol Specification, revision 2025-06-18 external_literature agent_tool_protocol Unassigned in current structure inter-stack-protocols-identity-and-economic-exchange source source note available; raw/cache text not published Official Model Context Protocol comparator for JSON-RPC message shape, lifecycle management, capability negotiation, session control, schema-defined interactions, and modular tool/client/server features; it does not establish a local protocol implementation, peer identity, authorization, message truth, task completion, payment, or deployment safety.
ext_a2a_protocol_0_3_0 Agent2Agent Protocol Specification, version 0.3.0 external_literature agent_to_agent_protocol Unassigned in current structure inter-stack-protocols-identity-and-economic-exchange source source note available; raw/cache text not published Official A2A comparator for agent discovery, agent cards, delegated tasks, artifact/message exchange, transport choices, and interoperability between opaque agent systems; it does not establish a local A2A deployment, verified identity, delegated authority, task truth, secure execution, payment, or safety.
ext_mcp_protocol_2025_11_25 Model Context Protocol Specification, revision 2025-11-25 external_literature agent_tool_protocol inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange) inter-stack-protocols-identity-and-economic-exchange source source note available; raw/cache text not published Latest released MCP comparator inspected on 2026-07-10 for versioned lifecycle, capability negotiation, authorization and OpenID Connect discovery changes, elicitation, tasks, and modular client/server boundaries; the announced 2026-07-28 revision remains a release candidate and is not represented as released.
ext_a2a_protocol_1_0_0 Agent2Agent Protocol Specification, version 1.0.0 external_literature agent_to_agent_protocol inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange) inter-stack-protocols-identity-and-economic-exchange source source note available; raw/cache text not published Latest released A2A comparator inspected on 2026-07-10 for canonical data objects, version negotiation, Agent Cards, tasks/messages/artifacts, JSON-RPC, gRPC and HTTP bindings, authorization scoping, interoperability testing, and security considerations; it does not establish local conformance, peer truth, delegated authority, or safe effects.
ext_w3c_did_core_1_0_2022 Decentralized Identifiers (DIDs) v1.0 external_literature decentralized_identity inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange) inter-stack-protocols-identity-and-economic-exchange source source note available; raw/cache text not published W3C DID Core comparator for decentralized identifier syntax, data model, controller-related metadata, resolution, and privacy considerations; it does not establish an ASI Stack identity system, controller trust, authorization, non-repudiation, revocation effectiveness, or safety.
ext_w3c_vc_data_model_2_0_2025 Verifiable Credentials Data Model v2.0 external_literature verifiable_credentials inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange) inter-stack-protocols-identity-and-economic-exchange source source note available; raw/cache text not published W3C Verifiable Credentials comparator for issuer/holder/verifier roles, credential and presentation fields, validity/status, evidence, securing mechanisms, and explicit authorization limitations; it does not establish an ASI Stack credential, trust decision, authorization framework, delegation validity, payment, or safety.
ext_interledger_protocol_v4 Interledger Protocol V4 external_literature interledger_value_transfer inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange) inter-stack-protocols-identity-and-economic-exchange source source note available; raw/cache text not published Official Interledger comparator for neutral packetized value transfer across independent ledgers, connector obligations, balances, and end-to-end boundary design; it does not establish an ASI Stack payment route, settlement, accounting correctness, legal transfer, economic fairness, delegated authority, or safety.
ext_test_time_compute_scaling_2024 Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters external_literature test_time_compute_allocation governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling) governed-deliberation-and-test-time-scaling source source note available; raw/cache text not published Primary test-time-compute comparator for verifier-guided search, proposal refinement, difficulty-dependent compute allocation, and the limits of extra inference; it does not establish local reasoning improvement, verifier correctness, safety, or an ASI Stack result.
ext_graphrag_2024 From Local to Global: A Graph RAG Approach to Query-Focused Summarization external_literature graph_based_retrieval_and_global_sensemaking virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) virtual-context-abi, context-transactions-snapshots-mounts-and-taint, verification-bandwidth-and-context-adequacy, open-research-agenda-and-bibliography-plan source source note available; raw/cache text not published Primary GraphRAG comparator for LLM-derived entity graphs, community summaries, and global corpus questions; generated graph and summary layers remain fallible derived representations and do not establish truth, complete coverage, local adequacy, or an ASI Stack memory result.
ext_hipporag_2024 HippoRAG: Neurobiologically Inspired Long-Term Memory for Large Language Models external_literature associative_long_term_memory virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) virtual-context-abi, verification-bandwidth-and-context-adequacy, routing-heads-and-specialist-cores, open-research-agenda-and-bibliography-plan source source note available; raw/cache text not published Primary NeurIPS comparator for knowledge-graph retrieval with Personalized PageRank and single-step associative navigation; reported multi-hop QA gains do not establish durable truth, update correctness, resistance to poisoning, local reproduction, or a general memory system.
ext_raptor_2024 RAPTOR: Recursive Abstractive Processing for Tree-Organized Retrieval external_literature hierarchical_retrieval_and_abstraction virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression) virtual-context-abi, verification-bandwidth-and-context-adequacy, compact-generative-systems-and-residual-honesty, rankfold-neuralfold-and-artifact-compression source source note available; raw/cache text not published Primary ICLR comparator for recursive clustering, summarization, and retrieval across multiple abstraction levels; source-reported QA gains do not prove summary fidelity, provenance preservation, local reproduction, or safe compression for ASI Stack claims.
ext_mem0_2025 Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory external_literature agent_long_term_memory virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) virtual-context-abi, context-transactions-snapshots-mounts-and-taint, procedural-memory-and-cognitive-loop-closure, resource-economics-and-token-budgets, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published Primary preprint comparator for extracting, consolidating, retrieving, and graph-linking conversational memory under latency and token-cost constraints; LOCOMO and LLM-judge results do not establish fact correctness, poisoning resistance, general memory, local reproduction, or production readiness here.
ext_w3c_prov_o_2013 PROV-O: The PROV Ontology external_literature interoperable_provenance_model evidence-states-and-claim-discipline (Evidence States and Claim Discipline); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) evidence-states-and-claim-discipline, claim-ledgers-and-belief-revision, artifact-graphs-audit-logs-and-replay, ai-supply-chain-integrity-and-lifecycle-provenance, data-engines-continual-learning-and-unlearning source source note available; raw/cache text not published W3C Recommendation comparator for interoperable provenance over entities, activities, agents, derivation, attribution, delegation, revision, and invalidation; a PROV-O graph records asserted provenance and does not by itself prove assertion truth, completeness, integrity, authority, or safety.
ext_mlcommons_croissant_1_1_2026 Croissant Format Specification, version 1.1 external_literature machine_readable_dataset_metadata ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) ai-supply-chain-integrity-and-lifecycle-provenance, artifact-graphs-audit-logs-and-replay, benchmark-ratchets-and-anti-goodhart-evidence, data-engines-continual-learning-and-unlearning source source note available; raw/cache text not published Current MLCommons specification comparator for JSON-LD dataset structure, resources, checksums, record fields, machine-readable provenance, usage conditions, and portability across ML tooling; metadata conformance does not prove dataset integrity, fitness, legality, representativeness, or safe use.
ext_inspect_ai_2024 Inspect AI: Framework for Large Language Model Evaluations external_literature model_and_agent_evaluation_framework runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); capability-thresholds-and-deployment-commitments (Capability Thresholds and Deployment Commitments); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) runtime-adapters-tool-permissions-and-human-approval, benchmark-ratchets-and-anti-goodhart-evidence, capability-thresholds-and-deployment-commitments, adversarial-evaluation-sandbagging-and-training-time-deception source source note available; raw/cache text not published Official UK AI Security Institute framework comparator for composable evaluation tasks, datasets, solvers, scorers, agents, tools, logs, and sandboxes; framework availability or a passing task does not establish benchmark validity, coverage, local execution, safety, or deployment readiness.
ext_in_toto_2019 in-toto: Providing farm-to-table guarantees for bits and bytes external_literature software_supply_chain_attestation model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay) model-weight-custody-and-hardware-roots-of-trust, ai-supply-chain-integrity-and-lifecycle-provenance, artifact-graphs-audit-logs-and-replay source source note available; raw/cache text not published Primary USENIX comparator for cryptographically verifying authorized software-supply-chain steps from source through deployment; valid attestations do not prove artifact correctness, uncompromised authorized actors, model safety, data fitness, or deployment merit.
ext_agentdojo_2024 AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents external_literature agent_prompt_injection_evaluation security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) security-kernel-and-digital-scifs, runtime-adapters-tool-permissions-and-human-approval, benchmark-ratchets-and-anti-goodhart-evidence, adversarial-evaluation-sandbagging-and-training-time-deception source source note available; raw/cache text not published Primary NeurIPS benchmark comparator for agents executing tools over untrusted data, with realistic tasks, security test cases, attacks, and defenses; benchmark results do not establish complete attack coverage, deployed robustness, safe authority handling, or local reproduction.
ext_camel_prompt_injection_2025 Defeating Prompt Injections by Design external_literature capability_secure_agent_control_flow system-boundaries-and-authority (System Boundaries and Authority); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) system-boundaries-and-authority, security-kernel-and-digital-scifs, intent-to-execution-contracts, runtime-adapters-tool-permissions-and-human-approval source source note available; raw/cache text not published Primary CaMeL comparator for separating trusted control flow from untrusted data and enforcing capability policies at tool calls; source-reported AgentDojo results do not establish universal prompt-injection resistance, correct policy extraction, local implementation, or safe deployment.
ext_owasp_agentic_top_10_2026 OWASP Top 10 for Agentic Applications for 2026 external_literature agentic_application_security_taxonomy system-boundaries-and-authority (System Boundaries and Authority); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) system-boundaries-and-authority, security-kernel-and-digital-scifs, ai-supply-chain-integrity-and-lifecycle-provenance, runtime-adapters-tool-permissions-and-human-approval, inter-stack-protocols-identity-and-economic-exchange, adversarial-evaluation-sandbagging-and-training-time-deception source source note available; raw/cache text not published Current OWASP community taxonomy comparator for goal hijacking, tool misuse, identity abuse, agentic supply chains, code execution, memory poisoning, inter-agent communication, cascading failures, human trust exploitation, and rogue agents; a risk list is not a proof of completeness, control effectiveness, local testing, or system safety.
ext_darwin_godel_machine_2025 Darwin Godel Machine: Open-Ended Evolution of Self-Improving Agents external_literature empirical_recursive_agent_improvement recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); open-ended-improvement-engines (Open-Ended Improvement Engines); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) recursive-self-improvement-boundaries, open-ended-improvement-engines, readiness-gates-residual-escrow-and-quarantine, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published Primary preprint comparator for archive-based open-ended code self-modification selected by empirical coding benchmarks under sandboxing and human oversight; reported benchmark gains do not establish monotonic general improvement, safe recursive self-improvement, local reproduction, or permission to self-modify.
ext_adas_2024 Automated Design of Agentic Systems external_literature automated_agent_architecture_search recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); open-ended-improvement-engines (Open-Ended Improvement Engines); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) recursive-self-improvement-boundaries, open-ended-improvement-engines, benchmark-ratchets-and-anti-goodhart-evidence, integrated-reference-architecture source source note available; raw/cache text not published Primary ADAS comparator for a meta-agent that searches a growing archive of code-defined agent designs across prompts, tools, and workflows; reported transfer results do not establish unrestricted generality, safe architecture search, local reproduction, or automatic promotion authority.
ext_universal_transformer_2019 Universal Transformers external_literature shared_weight_recurrence_and_adaptive_depth replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); mathematical-and-search-substrates (Mathematical and Search Substrates); coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts) governed-deliberation-and-test-time-scaling, mathematical-and-search-substrates, coil-attention-cyclic-memory-and-recurrence-contracts, replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary ICLR comparator for shared-weight depth recurrence, parallel self-attention, and per-position dynamic halting; benchmark results and theoretical expressivity do not establish stable deep recurrence, efficient scaling, local reproduction, or the book’s cyclic-memory claims.
ext_recurrent_transformer_2026 The Recurrent Transformer: Greater Effective Depth and Efficient Decoding external_literature layerwise_recurrent_transformer_memory fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets); mathematical-and-search-substrates (Mathematical and Search Substrates); coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts) fast-generation-architectures, resource-economics-and-token-budgets, mathematical-and-search-substrates, coil-attention-cyclic-memory-and-recurrence-contracts source source note available; raw/cache text not published Current preprint comparator for layerwise recurrent key-value memory, exact tiling, effective-depth/width tradeoffs, and standard autoregressive decoding cost; small-model C4 results do not establish broad capability gains, production efficiency, local reproduction, or cyclic-memory correctness.
ext_dynamic_compute_recurrent_transformers_2026 Understanding Dynamic Compute Allocation in Recurrent Transformers external_literature adaptive_recurrent_compute_evaluation replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); resource-economics-and-token-budgets (Resource Economics and Token Budgets); coil-attention-cyclic-memory-and-recurrence-contracts (Coil Attention, Cyclic Memory, and Recurrence Contracts); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) governed-deliberation-and-test-time-scaling, resource-economics-and-token-budgets, coil-attention-cyclic-memory-and-recurrence-contracts, benchmark-ratchets-and-anti-goodhart-evidence, replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Current preprint comparator for complexity-controlled tests of token-level variable-depth compute and online halting; its negative result that difficulty-aligned compute need not generalize is a boundary against equating adaptive depth with algorithmic extrapolation or local capability.
ext_claw_swe_bench_2026 Claw-SWE-Bench: A Benchmark for Evaluating OpenClaw-style Agent Harnesses on Coding Tasks external_literature coding_agent_harness_and_cost_evaluation artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) artifact-graphs-audit-logs-and-replay, runtime-adapters-tool-permissions-and-human-approval, resource-economics-and-token-budgets, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published Current preprint comparator for fixed workspace, patch, evaluator, and budget contracts across coding-agent harnesses; reported accuracy and cost differences are not reproduced here and do not validate the post-v2.1 synthetic repository corpus.
ext_txfs_2018 TxFS: Leveraging File-System Crash Consistency to Provide ACID Transactions external_literature transactional_filesystem_rollback_boundary capability-replacement-and-rollback (Capability Replacement and Rollback); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval) capability-replacement-and-rollback, artifact-graphs-audit-logs-and-replay, runtime-adapters-tool-permissions-and-human-approval source source note available; raw/cache text not published Primary systems comparator for user-space ACID file transactions built on journaling, including atomicity, isolation, durability, bounded transaction size, and Git/SQLite evaluation; it prevents treating a directory copy as a general transactional-filesystem result.
ext_dont_hallucinate_abstain_2024 Don’t Hallucinate, Abstain: Identifying LLM Knowledge Gaps via Multi-LLM Collaboration external_literature llm_abstention_and_knowledge_gaps verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine) routing-heads-and-specialist-cores, readiness-gates-residual-escrow-and-quarantine, verification-bandwidth-and-context-adequacy source source note available; raw/cache text not published Primary ACL comparator for knowledge-gap detection, abstention, calibration/self-reflection limitations, and multi-model probing; reported abstention improvements are task- and model-bounded and do not validate the local router or evaluator.
ext_muse_unlearning_2025 MUSE: Machine Unlearning Six-Way Evaluation for Language Models external_literature llm_unlearning_multidimensional_evaluation benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) data-engines-continual-learning-and-unlearning, policy-optimization-and-learning-from-feedback, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published Primary ICLR benchmark comparator separating verbatim and knowledge memorization, privacy leakage, retained utility, removal-scale behavior, and sequential sustainability; none of its 7B-language-model results are reproduced by the local policy network.
ext_unlearning_benchmarks_weak_2024 Position: LLM Unlearning Benchmarks are Weak Measures of Progress external_literature unlearning_benchmark_validity evidence-states-and-claim-discipline (Evidence States and Claim Discipline); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) data-engines-continual-learning-and-unlearning, benchmark-ratchets-and-anti-goodhart-evidence, evidence-states-and-claim-discipline source source note available; raw/cache text not published Primary critical comparator showing that benign benchmark modifications, forget/retain dependencies, and ambiguous targets can make unlearning scores optimistic; it strengthens the book’s prohibition on turning toy behavioral change into influence, privacy, or storage claims.
ext_openunlearning_2025 OpenUnlearning: Accelerating LLM Unlearning via Unified Benchmarking of Methods and Metrics external_literature unlearning_method_and_metric_meta_evaluation benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); living-book-methodology (Living Book Methodology) data-engines-continual-learning-and-unlearning, benchmark-ratchets-and-anti-goodhart-evidence, living-book-methodology source source note available; raw/cache text not published Primary NeurIPS 2025 benchmark-framework comparator for unified algorithms, diverse evaluations, public checkpoints, and meta-evaluation of metric faithfulness; it reinforces evaluator-quality residuals rather than establishing local unlearning.
qcsa_whitepaper Question-Compiled Semantic Addressing must_use semantic_addressing_control_plane governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture) cognitive-compilation-and-semantic-ir, virtual-context-abi, routing-heads-and-specialist-cores, compact-generative-systems-and-residual-honesty, runtime-adapters-tool-permissions-and-human-approval, claim-ledgers-and-belief-revision, data-engines-continual-learning-and-unlearning, inter-stack-protocols-identity-and-economic-exchange, integrated-reference-architecture, governed-world-models-and-reality-grounding, white-box-evidence-interpretability-and-activation-governance, durable-semantic-memory-and-knowledge-lattices source source note available; exact source published in the live-book paper library Corben-authored successor synthesis for stable semantic identity, plural versioned semantic virtual addresses, active question compilation, evidence-bearing hypergraphs, semantic address certificates, semantic-to-physical routing, lifecycle-safe migration, and explicit residuals. The later repository adds a bounded local 12-lane implementation, 60-case held-out evaluation over 13 systems and three seeds, and one 13-stage governed vertical trace. The matched-advantage and resource gates failed, the active-question ablation is N2 proxy/regime evidence rather than an exact or broad refutation, and no natural-task, learned-model, production, independent, chapter-core promotion, AGI, or ASI result is established.
reflexive_router_whitepaper The Reflexive Router: A Pre-Deliberative Architecture for Fast, Governed, Tool-Native Intelligence must_use pre_deliberative_reflexive_routing_control_plane stable-capability-fields (Stable Capability Fields); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture) routing-heads-and-specialist-cores, intent-to-execution-contracts, planning-as-a-control-layer, stable-capability-fields, virtual-context-abi, context-transactions-snapshots-mounts-and-taint, claim-ledgers-and-belief-revision, runtime-adapters-tool-permissions-and-human-approval, procedural-memory-and-cognitive-loop-closure, resource-economics-and-token-budgets, benchmark-ratchets-and-anti-goodhart-evidence, integrated-reference-architecture, white-box-evidence-interpretability-and-activation-governance source source note available; exact source published in the live-book paper library Corben-authored version 1.2 architecture proposal for a pre-deliberative command and routing plane, qualification-first dispatch, calibrated abstention, bounded execution DAGs, stable capability contracts, a non-bypassable effect commit kernel, typed result continuity, bitemporal Chronicle records, and governed trace-to-reflex compilation. It is assigned to existing chapter owners first; it adds no standalone chapter and supplies no implementation, benchmark, safety, deployment, transfer, novelty, AGI, ASI, or support-state result.
ext_faithfulness_information_flow_2026 Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning external_literature reasoning_trace_faithfulness artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) artifact-graphs-audit-logs-and-replay, adversarial-evaluation-sandbagging-and-training-time-deception, governed-deliberation-and-test-time-scaling, policy-optimization-and-learning-from-feedback source source note available; raw/cache text not published Primary 2026 comparator that separates chain-of-thought sufficiency, completeness, and interventional necessity, demonstrates prompt-to-answer shortcuts and transparent reward-hacking diagnostics, and documents low-entropy and reference-model limits. It does not make a reasoning transcript an authoritative receipt or establish local monitorability.
ext_monitorbench_2026 MonitorBench: A Comprehensive Benchmark for Chain-of-Thought Monitorability in Large Language Models external_literature reasoning_trace_monitorability_evaluation scalable-oversight-and-adversarial-ai-control (Scalable Oversight and Adversarial AI Control); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception) scalable-oversight-and-adversarial-ai-control, adversarial-evaluation-sandbagging-and-training-time-deception source source note available; raw/cache text not published Primary open benchmark comparator with 1,514 instances across 19 tasks and seven categories plus two stress-test settings; its reported capability/monitorability relation and up-to-30-percent degradation motivate held-out trace-action stress tests. The benchmark does not establish local monitoring quality, causal trace faithfulness, or safety.
ext_v_jepa_2_2025 V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning external_literature latent_world_models_and_model_predictive_control planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); mathematical-and-search-substrates (Mathematical and Search Substrates); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture) mathematical-and-search-substrates, planning-as-a-control-layer, data-engines-continual-learning-and-unlearning, governed-world-models-and-reality-grounding, integrated-reference-architecture source source note available; raw/cache text not published Primary empirical comparator for action-free latent video pretraining, a small action-conditioned predictor, and model-predictive control. Camera sensitivity, autoregressive error accumulation, action-search cost, image-goal assumptions, and representation-bounded capability remain explicit limits; no local world model or robot-control result is established.
ext_embedded_agency_2019 Embedded Agency external_literature embedded_agency_foundations asi-is-a-stack-not-a-model (ASI Is a Stack, Not a Model); evidence-states-and-claim-discipline (Evidence States and Claim Discipline); constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); integrated-reference-architecture (Integrated Reference Architecture) asi-is-a-stack-not-a-model, constitutional-alignment-substrate, recursive-self-improvement-boundaries, evidence-states-and-claim-discipline, integrated-reference-architecture source source note available; raw/cache text not published Primary informal survey of the obstacles that arise when agents are physical parts of the worlds they model, must use smaller internal models, and reason about modifiable internal parts. It supplies a foundations boundary; the book’s finite records, authority ceilings, and proofs do not solve embedded agency.
ext_ietf_rats_architecture_2023 Remote ATtestation procedureS (RATS) Architecture external_literature remote_attestation_architecture confidential-and-verifiable-ai-computation (Confidential and Verifiable AI Computation); model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust) model-weight-custody-and-hardware-roots-of-trust source source note available; raw/cache text not published Primary IETF architecture and terminology comparator for Attester, Verifier, Relying Party, Evidence, Attestation Results, appraisal policies, reference values, freshness, layered environments, privacy, trust roots, and confidential-model key release. It is informational architecture, not a protocol, hardware assurance level, verifier-independence result, or local attestation deployment.
ext_nist_key_management_2020 Recommendation for Key Management: Part 1 – General external_literature cryptographic_key_management model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust) model-weight-custody-and-hardware-roots-of-trust source source note available; raw/cache text not published Current final NIST key-management baseline for key and metadata protection, inventory, authorization, access control, usage periods, compromise, backup, recovery, trust anchors, and lifecycle policy. A Revision 6 draft exists, so this source is the final baseline rather than a claim that guidance has stopped evolving; no local key-management conformance or security result is established.
ext_nist_media_sanitization_2025 Guidelines for Media Sanitization external_literature media_sanitization_and_disposal model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust) model-weight-custody-and-hardware-roots-of-trust source source note available; raw/cache text not published Current final NIST media-sanitization comparator for rendering target data access infeasible at a stated effort level using sensitivity- and media-appropriate controls, including cryptographic erase. It does not prove that all model copies, plaintext memory, cloud replicas, derivatives, or recipients were discovered or sanitized, and no local erasure test was run.
corben_chatgpt_kiss_irreducible_intelligence_2026 KISS versus Irreducible Intelligence (author-supplied design conversation) supporting author_intent_replaceable_cognitive_substrates replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture, relational-dimension-compilation-and-polyadic-cognition source source note available; raw/cache text not published Corben-supplied design conversation for search-verify-compile, exact-latent separation, a compact recursive kernel, and total-system KISS accounting. Author intent only; not independent evidence or a reproduced architecture.
corben_chatgpt_onecell_theseus_2026 OneCell and Theseus Architecture Handoff (author-supplied design conversation) supporting author_intent_onecell_and_architectural_rsi replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Corben-supplied design conversation for a Cognitive Kernel ABI, typed state lanes, inner recurrence, outer exact search, verified abstraction, and Theseus-governed architecture tournaments. Author intent only; OneCell remains an unimplemented falsifiable candidate.
ext_attention_is_all_you_need_2017 Attention Is All You Need external_literature dense_attention_sequence_substrate replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary Transformer paper and dense-attention baseline. It supports the historical architecture and parallel sequence-processing comparison, not a claim that Transformers are universally optimal or locally reproduced.
ext_mamba2_ssd_2024 Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality external_literature state_space_duality_and_sequence_substrates replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary Mamba-2/structured-state-space-duality comparator connecting SSM and attention-like formulations. No local model, kernel, quality, scaling, or hardware result is reproduced.
ext_s4_2022 Efficiently Modeling Long Sequences with Structured State Spaces external_literature structured_state_space_sequence_substrates replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Foundational S4 comparator for structured state-space sequence models, efficient long-range computation, and the lineage that precedes selective SSMs. Source-reported benchmark and generation results are not reproduced and do not establish exact recall or governed substitutability.
ext_mamba3_2026 Mamba-3: Improved Sequence Modeling using State Space Principles external_literature modern_selective_state_space_sequence_substrates replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Current 2026 selective-SSM comparator for complex-valued state updates, discretization, and multi-input/multi-output formulation. Recent source-reported results are not locally reproduced and must not set the chapter conclusion by recency.
ext_gated_deltanet2_2026 Gated DeltaNet-2: Decoupling Erase and Write in Linear Attention external_literature current_recurrent_linear_attention_and_editable_memory_frontier replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Dated 2026 comparator that separates erase and write gates and reports the strongest aggregate result among its Mamba-2, Gated DeltaNet, KDA, Mamba-3, and Gated DeltaNet-2 envelope at 1.3B parameters and 100B FineWeb-Edu tokens. The result is author reported, not locally reproduced; it displaces Mamba-3 only for that exact source envelope and requires official-code, checkpoint, hardware, seed, cost, retrieval, state, and transfer reproduction before any local superiority claim.
ext_hyperscale_lottery_2026 The Hyperscale Lottery: How State-Space Models Have Sacrificed Edge Efficiency external_literature hardware_specific_state_space_efficiency_counterevidence replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Current edge-hardware counterstudy measuring Mamba-family latency outside hyperscale-GPU conditions. Its source-reported results require platform-stratified latency, memory, and energy accounting; they are not locally reproduced and do not settle the quality-efficiency frontier.
ext_gated_deltanet_2024 Gated Delta Networks: Improving Mamba2 with Delta Rule external_literature linear_attention_and_adaptive_memory replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary gated-delta-rule comparator for targeted recurrent-memory updates, rapid erasure, parallel training, and hybrid attention/SSM compositions. Source-reported retrieval, extrapolation, efficiency, and quality results are not locally reproduced.
ext_jamba_2024 Jamba: A Hybrid Transformer-Mamba Language Model external_literature hybrid_attention_state_space_mixture_of_experts replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary large-scale hybrid Transformer-Mamba-MoE comparator. It motivates route- and composition-aware accounting; its reported quality, context, throughput, and memory results are not locally reproduced and do not establish that the specific mixture is generally optimal.
ext_neural_message_passing_2017 Neural Message Passing for Quantum Chemistry external_literature graph_relational_message_passing replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); relational-dimension-compilation-and-polyadic-cognition (Relational Dimension Compilation and Polyadic Cognition) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary message-passing neural-network framework for learned computation over graph structure. Its molecular results motivate a non-token-native relational lane but do not establish general reasoning, dynamic graph memory, exact state, or local reproduction.
ext_hyena_hierarchy_2023 Hyena Hierarchy: Towards Larger Convolutional Language Models external_literature long_convolution_sequence_substrates replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary long-convolution comparator for subquadratic sequence mixing and hardware-aware architecture comparisons. No local training, throughput, quality, recall, or scaling result is reproduced.
ext_rwkv_2023 RWKV: Reinventing RNNs for the Transformer Era external_literature linear_recurrent_language_models replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary recurrent language-model comparator combining parallelizable training with recurrent inference. Reported benchmark, memory, and inference properties are not reproduced locally.
ext_xlstm_2024 xLSTM: Extended Long Short-Term Memory external_literature modern_gated_recurrent_sequence_substrates replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary modern-LSTM comparator for revised gating, memory, and scalable recurrent language modeling. No xLSTM training, scaling, quality, or inference result is reproduced locally.
ext_ttt_layers_2024 Learning to (Learn at Test Time): RNNs with Expressive Hidden States external_literature test_time_learned_state_sequence_substrates replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary test-time-training-layer comparator that treats hidden state as a learned model updated on the sequence. It motivates explicit online-state custody and rollback; no local quality or efficiency result is reproduced.
ext_titans_2025 Titans: Learning to Memorize at Test Time external_literature test_time_neural_long_term_memory durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary neural-memory comparator for test-time memorization and long-context sequence modeling. The paper motivates mutable-state provenance and rollback tests; no local model or benchmark result is reproduced.
ext_kan_2024 KAN: Kolmogorov-Arnold Networks external_literature learned_univariate_function_networks replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary KAN proposal replacing fixed node activations/linear edge weights with learned univariate edge functions. Interpretability and scientific-task demonstrations are source-reported and do not establish a general MLP or Transformer replacement.
ext_kan_or_mlp_fairer_comparison_2024 KAN or MLP: A Fairer Comparison external_literature architecture_comparison_methodology replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Critical matched-comparison source for KAN versus MLP under parameter, FLOP, and task controls. It is included to prevent architecture enthusiasm from substituting for fair accounting; no local comparison is reproduced.
ext_neural_turing_machines_2014 Neural Turing Machines external_literature differentiable_external_memory replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary differentiable-controller/external-memory source. It motivates variable-size memory interfaces and out-of-distribution algorithmic tests; toy-task results do not establish reliable exact memory or general computation.
ext_differentiable_neural_computer_2016 Hybrid computing using a neural network with dynamic external memory external_literature differentiable_external_memory replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary Differentiable Neural Computer source for learned controllers over dynamic external memory. Source-reported graph and reasoning tasks do not establish reliable exact state, scalable memory, or local reproduction.
ext_liquid_time_constant_networks_2021 Liquid Time-constant Networks external_literature continuous_time_neural_dynamics replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary continuous-time recurrent architecture source for input-dependent time constants and dynamical-system behavior. Reported time-series results and stability analysis do not establish broad cognitive superiority or a local implementation.
ext_tiny_recursive_model_2025 Less is More: Recursive Reasoning with Tiny Networks external_literature tiny_weight_tied_recursive_reasoning replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary Tiny Recursive Model proposal and narrow puzzle-domain result. It motivates a compact weight-tied OneCell comparator but does not establish general reasoning, language capability, deep effective recursion, or total-system simplicity.
ext_trm_arc_agi_analysis_2025 Tiny Recursive Models on ARC-AGI-1: Inductive Biases, Identity Conditioning, and Test-Time Compute external_literature recursive_model_critical_evaluation replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Critical TRM analysis reporting material dependence on 1000-sample voting, puzzle identity, and shallow effective recursion. It is a source-reported audit rather than a local reproduction and sets preregistered identity, sampling, and recursion-depth controls.
ext_tiny_autoregressive_recursive_models_2026 Tiny Autoregressive Recursive Models external_literature recursive_model_controlled_ablation replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Controlled compute-matched study that progressively transforms a standard autoregressive model into a TRM-like system and reports no reliable advantage from the full autoregressive TRM mechanism. It motivates mechanism-level rather than label-level ablation.
ext_unimatrix_2026 Associative-State Universal Transformers: Sparse Retrieval Meets Structured Recurrence external_literature structured_recurrence_and_sparse_retrieval replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Current UniMatrix preprint whose negative associative-recall result shows compressed recurrent state alone is insufficient in its setup, while explicit sparse slots and pointer-level routing materially change the result. No local reproduction or general conclusion follows.
ext_memory_caching_2026 Memory Caching: RNNs with Growing Memory external_literature growing_recurrent_memory replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Current recurrent-memory comparator that caches hidden-state checkpoints and exposes a trade between fixed recurrent memory and growing addressable memory. Its source-reported recall results still leave the Transformer strongest on the reported in-context recall tasks.
ext_inkling_2026 Inkling: Our open-weights model external_literature hybrid_local_global_attention_moe_multimodal_substrate replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Release-day primary-source case study of a 66-layer multimodal sparse-MoE Transformer with a 5:1 local/global attention schedule, relative positions, short convolutions, controllable effort, and open weights. It motivates topology-complete capability cards and component ablations; provider-reported results are not locally reproduced and do not isolate the contribution of any component.
kernel_english_residual_compiler Kernel English with Hierarchical, Interaction-Amortized Residuals: A Dual-Vocabulary Cognitive Compiler for Efficient Language-Model Reasoning must_use canonical_cognitive_compilation_and_hierarchical_residual_runtime security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); verification-bandwidth-and-context-adequacy (Verification Bandwidth and Context Adequacy); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance); integrated-reference-architecture (Integrated Reference Architecture) cognitive-compilation-and-semantic-ir, compact-generative-systems-and-residual-honesty, virtual-context-abi, context-transactions-snapshots-mounts-and-taint, verification-bandwidth-and-context-adequacy, fast-generation-architectures, replaceable-cognitive-substrates-beyond-transformer-monoculture, resource-economics-and-token-budgets, security-kernel-and-digital-scifs, procedural-memory-and-cognitive-loop-closure, benchmark-ratchets-and-anti-goodhart-evidence, integrated-reference-architecture, white-box-evidence-interpretability-and-activation-governance source source note available; exact source published in the live-book paper library Corben-authored July 2026 architecture proposal for KERC: protected-object capture, uncertainty-aware normalization, sense-aware Kernel IR, dual surface/core vocabularies, a four-level interaction-amortized residual ledger, exact object storage, grammar-aware macro fusion, structured answer packets, rendering, round-trip verification, versioned migration, and complete rate-compute-fidelity evaluation. Existing chapters are upgraded first; no implementation, benchmark, novelty, efficiency, fidelity, safety, transfer, SOTA, AGI, ASI, or support-state result is inferred.
deterministic_capability_compilation Deterministic Capability Compilation: A Capability-Preserving Ladder from Executable Scaffolds to Governed Adaptive Agents must_use capability_compilation_neural_linking_and_governed_adaptation stable-capability-fields (Stable Capability Fields); capability-replacement-and-rollback (Capability Replacement and Rollback); adversarial-machine-learning-and-model-attack-surface (Adversarial Machine Learning and the Model Attack Surface); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); autonomous-replication-proliferation-and-containment (Autonomous Replication, Proliferation, and Containment); intent-to-execution-contracts (Command Contracts: From Intent to Executable Work); governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); labor-os-and-typed-jobs (Labor OS and Typed Jobs); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance); content-authenticity-watermarking-and-synthetic-media-integrity (Content Authenticity, Watermarking, and Synthetic Media Integrity); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture); project-theseus-as-report-first-implementation-reference (Project Theseus as Report-First Implementation Reference); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) stable-capability-fields, intent-to-execution-contracts, cognitive-compilation-and-semantic-ir, virtual-context-abi, capability-replacement-and-rollback, routing-heads-and-specialist-cores, runtime-adapters-tool-permissions-and-human-approval, spinoza-verification-and-proof-carrying-claims, labor-os-and-typed-jobs, artifact-graphs-audit-logs-and-replay, procedural-memory-and-cognitive-loop-closure, readiness-gates-residual-escrow-and-quarantine, compact-generative-systems-and-residual-honesty, replaceable-cognitive-substrates-beyond-transformer-monoculture, ai-supply-chain-integrity-and-lifecycle-provenance, recursive-self-improvement-boundaries, benchmark-ratchets-and-anti-goodhart-evidence, data-engines-continual-learning-and-unlearning, integrated-reference-architecture, project-theseus-as-report-first-implementation-reference, prototype-roadmap, open-research-agenda-and-bibliography-plan, white-box-evidence-interpretability-and-activation-governance, governed-world-models-and-reality-grounding, governed-operations-incident-command-and-graceful-degradation, adversarial-machine-learning-and-model-attack-surface, autonomous-replication-proliferation-and-containment source source note available; exact source published in the live-book paper library Corben-authored July 2026 architecture and research program for compiling executable scaffolds into contract-bound experts and linked Neural Capability Objects while retaining semantic obligation mass balance, candidate-specific translation validation, fallback, residual escrow, authority ceilings, reification, and effect-complete recovery. Existing chapters are upgraded first; no foundry implementation, learned-capability result, preservation result, safety result, SOTA result, AGI, ASI, or support-state promotion is inferred.
platonic_world_model The Platonic World Model: A Semantic Constitution for Grounded, Proof-Carrying, Self-Editing Artificial Intelligence must_use semantic_continuity_grounded_world_model_and_governed_self_editing moral-uncertainty-and-value-conflict (Moral Uncertainty, Value Conflict, and Contestable Governance); security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); perception-sensor-fusion-and-observation-trust (Perception, Sensor Fusion, and Observation Trust); planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); spinoza-verification-and-proof-carrying-claims (Proof-Carrying Claims and Adversarial Review); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture); prototype-roadmap (Prototype Roadmap); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) moral-uncertainty-and-value-conflict, security-kernel-and-digital-scifs, ai-supply-chain-integrity-and-lifecycle-provenance, claim-ledgers-and-belief-revision, spinoza-verification-and-proof-carrying-claims, artifact-graphs-audit-logs-and-replay, virtual-context-abi, context-transactions-snapshots-mounts-and-taint, planning-as-a-control-layer, cognitive-compilation-and-semantic-ir, runtime-adapters-tool-permissions-and-human-approval, inter-stack-protocols-identity-and-economic-exchange, procedural-memory-and-cognitive-loop-closure, data-engines-continual-learning-and-unlearning, recursive-self-improvement-boundaries, benchmark-ratchets-and-anti-goodhart-evidence, integrated-reference-architecture, prototype-roadmap, open-research-agenda-and-bibliography-plan, white-box-evidence-interpretability-and-activation-governance, governed-world-models-and-reality-grounding, governed-operations-incident-command-and-graceful-degradation source source note available; exact source published in the live-book paper library Corben-authored July 2026 conceptual architecture and falsifiable research program for semantic continuity through stable Form lineages, immutable semantic versions, typed Essence Contracts, six mutually constraining planes, explicit proposition-attestation-commitment-proof separation, branch-protected world dynamics, qualified grounding, semantic transactions, runtime packet compilation, and federated mappings. Existing chapters are upgraded first; no implemented substrate, benchmark result, philosophical solution to grounding, safety result, SOTA result, AGI, ASI, or support-state promotion is inferred.
relational_dimension_compiler The Relational Dimension Compiler: Adaptive Polyadic Cognition with Bounded Computational Arity and Unbounded Semantic Structure must_use typed_relational_ir_adaptive_polyadic_routing_and_reversible_abstraction governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); relational-dimension-compilation-and-polyadic-cognition (Relational Dimension Compilation and Polyadic Cognition); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); integrated-reference-architecture (Integrated Reference Architecture); open-research-agenda-and-bibliography-plan (Open Research Agenda and Bibliography Plan) cognitive-compilation-and-semantic-ir, governed-world-models-and-reality-grounding, routing-heads-and-specialist-cores, replaceable-cognitive-substrates-beyond-transformer-monoculture, procedural-memory-and-cognitive-loop-closure, benchmark-ratchets-and-anti-goodhart-evidence, resource-economics-and-token-budgets, integrated-reference-architecture, open-research-agenda-and-bibliography-plan source source note available; exact source published in the live-book paper library Corben-authored July 2026 conceptual architecture and falsifiable research program for a typed relational intermediate representation that separates geometric dimension, semantic arity, primitive computational arity, storage arity, temporal extent, abstraction scale, branch identity, epistemic status, and resource budget. It proposes sparse adaptive relational-order routing, exact role-preserving relation reification, qualified relation lifecycles, branch-local object-field state, reversible semantic contraction, compiled slow-to-fast relation programs, hardware lowering, and the RODIE benchmark suite. Existing chapters receive bounded integration while a distinct future chapter candidate remains deferred by the active manifest freeze. No RDC implementation, benchmark result, universal arity bound, ontology truth, efficiency, safety, SOTA, AGI, ASI, or support-state promotion is inferred.
ext_megatron_distributed_training_2021 Efficient Large-Scale Language Model Training on GPU Clusters Using Megatron-LM external_literature governed_distributed_model_training_and_scaling governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); learning-compute-topology-and-adaptive-process-architecture (Learning–Compute Topology and Adaptive Process Architecture) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary composed-parallelism mechanism source. It grounds tensor, pipeline, and data parallel interactions, strict optimizer semantics, microbatch and topology tradeoffs. Reported trillion-parameter and throughput results are configuration-bound and not locally reproduced.
ext_zero_optimizer_2019 ZeRO: Memory Optimizations Toward Training Trillion Parameter Models external_literature governed_distributed_model_training_and_scaling governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary competing sharded-state design for optimizer, gradient, parameter, activation, and residual memory. It motivates explicit state closure and shard reconstruction; source-reported scale and speed are not locally reproduced or treated as universal superiority.
ext_gspmd_2021 GSPMD: General and Scalable Parallelization for ML Computation Graphs external_literature governed_distributed_model_training_and_scaling governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); learning-compute-topology-and-adaptive-process-architecture (Learning–Compute Topology and Adaptive Process Architecture) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary compiler-mediated competing design for general SPMD sharding and mixed parallelism. It motivates versioning inferred plans and inserted collectives; reported TPU utilization and scaling are not locally reproduced.
ext_datastates_llm_2024 DataStates-LLM: Lazy Asynchronous Checkpointing for Large Language Models external_literature governed_distributed_model_training_and_scaling governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary limitation and checkpoint-mechanism source for asynchronous multi-level copies, distributed shard consistency, and checkpoint overhead. It does not establish complete application state or exact trajectory-equivalent resume, and no result is locally reproduced.
ext_pytorch_distributed_checkpoint_2026 Distributed Checkpoint — PyTorch documentation external_literature governed_distributed_model_training_and_scaling governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Official current implementation documentation for SPMD save/load, asynchronous completion, canonical model and optimizer state, resharding, strict load, and call-order constraints. Documentation is not benchmark or full-state resume evidence.
ext_mlperf_training_v6_2026 MLPerf Training v6.0 external_literature governed_distributed_model_training_and_scaling governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) governed-model-training-distributed-optimization-and-scaling, benchmark-ratchets-and-anti-goodhart-evidence, resource-economics-and-token-budgets source source note available; raw/cache text not published Official current measurement comparator for fixed datasets and quality targets, repeated time-to-quality, system metadata, divisions, variance, and corrected results. No MLPerf run is performed and the benchmark does not establish safety or complete run integrity.
ext_adam_2015 Adam: A Method for Stochastic Optimization external_literature optimizer_mechanisms_and_selection governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary Adam mechanism source for bias-corrected first- and second-moment estimates and coordinate-wise adaptive updates. Its online-convex analysis and reported experiments do not establish universal convergence, quality, or optimizer superiority in foundation-model training.
ext_amsgrad_2018 On the Convergence of Adam and Beyond external_literature optimizer_failure_and_convergence governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary Adam failure and AMSGrad source. It gives constructed stochastic-convex non-convergence cases and a maximum-second-moment remedy; those cases do not imply every practical Adam run fails or that AMSGrad is universally preferable.
ext_adamw_2019 Decoupled Weight Decay Regularization external_literature optimizer_mechanisms_and_selection governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary AdamW source separating weight decay from the adaptive gradient update. Its reported tuning and generalization results are setting-bound; the optimizer name alone does not specify parameter exclusions, schedule, decay scaling, or implementation semantics.
ext_adafactor_2018 Adafactor: Adaptive Learning Rates with Sublinear Memory Cost external_literature optimizer_memory_and_scaling governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary factored-second-moment optimizer source. It reduces auxiliary state for matrix parameters and adds update clipping and parameter-scale rules; factorization remains an approximation and its reported translation result does not establish universal parity with Adam.
ext_lamb_2019 Large Batch Optimization for Deep Learning: Training BERT in 76 minutes external_literature optimizer_memory_and_scaling governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary LAMB source for layer-wise trust ratios and large-batch optimization. Its reported BERT time-to-target is tied to model, batch, hardware, quality target, and tuning conditions and is not a universal large-batch or wall-clock result.
ext_shampoo_2018 Shampoo: Preconditioned Stochastic Tensor Optimization external_literature tensor_and_matrix_preconditioning governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary tensor-structured preconditioning source. Shampoo maintains per-dimension preconditioners and reports faster convergence with practical per-step cost in studied models; stochastic-convex theory and source experiments do not settle current distributed lifecycle cost.
ext_kfac_2015 Optimizing Neural Networks with Kronecker-factored Approximate Curvature external_literature curvature_aware_optimization governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary K-FAC source for an efficiently invertible Kronecker-factored approximation to the Fisher matrix. Its approximation, damping, inversion, and empirical cost-benefit are architecture- and implementation-dependent.
ext_lion_2023 Symbolic Discovery of Optimization Algorithms external_literature optimizer_discovery_and_sign_updates governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary Lion and symbolic optimizer-search source. Lion uses sign-based momentum and one optimizer-state tensor; the paper also reports method-specific learning-rate behavior and settings where gains are small or insignificant.
ext_sophia_2023 Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training external_literature curvature_aware_optimization governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary Sophia source for periodic diagonal-Hessian estimation and clipped curvature-aware updates. Its reported GPT pretraining speedups and simplified theory require matched reproduction before any broader optimizer claim.
ext_soap_2024 SOAP: Improving and Stabilizing Shampoo using Adam external_literature tensor_and_matrix_preconditioning governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary SOAP source connecting Shampoo to adaptive moments in a changing preconditioner eigenbasis. Reported large-batch pretraining gains remain tied to 360M/660M models, preconditioning frequency, overhead, and tuning conditions.
ext_schedule_free_2024 The Road Less Scheduled external_literature optimizer_scheduling_and_averaging governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary schedule-free optimization source unifying scheduling and iterate averaging without requiring a stopping step. Removing a stopping-time schedule does not remove learning-rate, warmup, evaluation-iterate, checkpoint, or method-selection choices.
ext_mup_2022 Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer external_literature optimizer_parametrization_and_scale_transfer governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary maximal-update parametrization and muTransfer source. It reports widthwise hyperparameter transfer under a prescribed parametrization on Transformer and ResNet settings; it does not establish arbitrary depth, duration, optimizer, or architecture transfer.
ext_modular_norm_2024 Scalable Optimization in the Modular Norm external_literature optimizer_parametrization_and_scale_transfer governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary modular-norm source defining architecture-recursive update geometry and reporting learning-rate transfer across width and depth. Its well-behaved-module assumptions and experiments do not prove arbitrary substrate transfer or optimizer superiority.
ext_muon_scalable_2025 Muon is Scalable for LLM Training external_literature orthogonalized_matrix_optimization governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary large-scale Muon source for momentum plus matrix orthogonalization, weight decay, per-parameter update scaling, and a distributed implementation. Its reported compute-efficiency and Moonlight results are source-scoped and not locally reproduced.
ext_muon_spectral_norm_2026 Muon Optimizes Under Spectral Norm Constraints external_literature orthogonalized_matrix_optimization_theory governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Accepted TMLR theoretical source placing Muon with decoupled weight decay in a Lion-K/nuclear-norm framework and deriving implicit spectral-norm constraint behavior. The interpretation does not establish task-level quality, safety, or universal advantage.
ext_nist_privacy_framework_2020 NIST Privacy Framework: A Tool for Improving Privacy through Enterprise Risk Management, Version 1.0 external_literature privacy_data_rights_and_information_flow_governance privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) privacy-data-rights-and-information-flow-governance, security-kernel-and-digital-scifs source source note available; raw/cache text not published Official paper-body-reviewed risk framework distinguishing privacy problems from cybersecurity incidents across the data lifecycle. It is voluntary, has no force of law, and supplies no local privacy outcome or certification.
ext_eu_gdpr_2016 Regulation (EU) 2016/679 (General Data Protection Regulation) external_literature privacy_data_rights_and_information_flow_governance privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) privacy-data-rights-and-information-flow-governance source source note available; raw/cache text not published Authoritative jurisdiction-specific normative comparator for principles, bases, rights, accountability, design, and qualified exceptions. It is not universal law, legal advice, an applicability decision, or local compliance evidence.
ext_w3c_dpv_2024 Data Privacy Vocabulary (DPV), Version 2 external_literature privacy_data_rights_and_information_flow_governance privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) privacy-data-rights-and-information-flow-governance, context-transactions-snapshots-mounts-and-taint source source note available; raw/cache text not published Machine-readable vocabulary for purpose, processing, data, actors, rights, risks, measures, legal basis, and consent. It is a Community Group Final Specification, not a W3C Recommendation, law, or enforcement proof.
ext_abadi_dpsgd_2016 Deep Learning with Differential Privacy external_literature privacy_data_rights_and_information_flow_governance privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) privacy-data-rights-and-information-flow-governance, governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary DP-SGD mechanism and accounting source. Its algorithm, analysis, and reported experiments are not locally reproduced; its guarantee is parameter-, unit-, adjacency-, implementation-, and release-surface-bound.
ext_algospec_purpose_limitation_2024 Being Transparent Is Merely the Beginning: Enforcing Purpose Limitation with Polynomial Approximation external_literature privacy_data_rights_and_information_flow_governance privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) privacy-data-rights-and-information-flow-governance source source note available; raw/cache text not published Primary competing purpose-restriction design using algorithm-specific polynomial approximation. Reported accuracy and efficiency are bounded to studied algorithms/data and are not locally reproduced or a complete legal-purpose result.
ext_carlini_training_data_extraction_2021 Extracting Training Data from Large Language Models external_literature privacy_data_rights_and_information_flow_governance adversarial-machine-learning-and-model-attack-surface (Adversarial Machine Learning and the Model Attack Surface); privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) privacy-data-rights-and-information-flow-governance, data-engines-continual-learning-and-unlearning source source note available; raw/cache text not published Primary failure source reporting black-box extraction of memorized GPT-2 training sequences. The source result is configuration-bound and not a local or universal leakage result.
ext_choquette_choo_label_only_mia_2021 Label-Only Membership Inference Attacks external_literature privacy_data_rights_and_information_flow_governance privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) privacy-data-rights-and-information-flow-governance source source note available; raw/cache text not published Primary failure source showing hard-label robustness can expose membership and confidence masking can be insufficient in studied settings. No attack or defense result is locally reproduced or universal.
ext_mahloujifar_fdp_audit_2025 Auditing f-Differential Privacy in One Run external_literature privacy_data_rights_and_information_flow_governance privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance) privacy-data-rights-and-information-flow-governance, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published Primary empirical-audit comparator using randomized inclusion and an f-DP hypothesis in one run. A passed audit is not proof that DP or lifecycle privacy holds, and no result is locally reproduced.
ext_airllm_2023 AirLLM: Scaling Large Language Models on Low-End Commodity Computers external_literature heterogeneous_inference_memory model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets, model-weight-custody-and-hardware-roots-of-trust source source note available; raw/cache text not published Official implementation comparator for layer-wise model sharding, one-layer accelerator residency, next-layer prefetch, optional storage compression, and original-versus-transformed model storage. Maintainer-reported fit and speed claims are not independently reproduced.
ext_deepspeed_inference_2022 DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale external_literature heterogeneous_inference_memory personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets source source note available; raw/cache text not published Primary heterogeneous-inference systems source spanning GPU, CPU, and NVMe for dense and sparse Transformer inference. Reported latency, throughput, scale, and model-fit results remain source-scoped and unreproduced.
ext_flexgen_2023 FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU external_literature heterogeneous_inference_memory personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets source source note available; raw/cache text not published Primary planned-placement source for GPU/CPU/disk tensor storage and access, batching, and optional weight/cache compression under latency-insensitive workloads. Its throughput results are not interactive-latency or local evidence.
ext_hf_accelerate_big_model_inference_2026 Hugging Face Accelerate: Loading Big Models into Memory external_literature heterogeneous_inference_memory model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); virtual-context-abi (The Virtual Context ABI: Typed Pages, Cells, and Certificates); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets, model-weight-custody-and-hardware-roots-of-trust source source note available; raw/cache text not published Official implementation documentation for automatic or explicit GPU/CPU/disk device maps and memory-mapped disk tensors. The documented sequential-dispatch, prefetch, and hard-drive-performance limitations make it a baseline, not a qualification result.
ext_llama_cpp_memory_mapping_2026 llama.cpp CLI Memory Mapping, Tensor Placement, and KV Offload Controls external_literature heterogeneous_inference_memory model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets, model-weight-custody-and-hardware-roots-of-trust source source note available; raw/cache text not published Official consumer-runtime documentation for model load modes, memory mapping, DirectIO, GPU-layer and tensor placement, MoE CPU placement, KV offload, and KV data types. No local model or performance result is implied.
ext_llm_in_flash_2024 LLM in a Flash: Efficient Large Language Model Inference with Limited Memory external_literature heterogeneous_inference_memory model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets, model-weight-custody-and-hardware-roots-of-trust, replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary flash-aware inference source for on-demand parameter loading, I/O cost modeling, transfer reduction, contiguous reads, windowing, and row-column bundling. Sparse/context-adaptive loading is not an exact dense paging result.
ext_powerinfer_2024 PowerInfer: Fast Large Language Model Serving with a Consumer-Grade GPU external_literature heterogeneous_inference_memory replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets, replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Primary consumer-inference source for source-reported power-law neuron locality, hot-GPU/cold-CPU placement, adaptive predictors, and sparse operators. Architecture transfer and performance are not locally reproduced.
ext_vattention_2025 vAttention: Dynamic Memory Management for Serving LLMs without PagedAttention external_literature heterogeneous_inference_memory fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, resource-economics-and-token-budgets source source note available; raw/cache text not published Primary counterpoint to non-contiguous PagedAttention layouts: decouples virtual and physical GPU memory while retaining contiguous KV virtual addresses. Reported serving results remain source-scoped.
ext_infinigen_2024 InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache Management external_literature heterogeneous_inference_memory fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, resource-economics-and-token-budgets source source note available; raw/cache text not published Primary speculative-KV-prefetch source using minimal rehearsal and partial next-layer state to select host-resident KV entries. Prediction, quality, miss, and fallback results are not locally reproduced.
ext_specache_2025 SpeCache: Speculative Key-Value Caching for Efficient Generation of LLMs external_literature heterogeneous_inference_memory fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, resource-economics-and-token-budgets source source note available; raw/cache text not published Primary speculative-KV-prefetch source keeping complete KV state in CPU memory, a low-bit importance copy in VRAM, and predicted next-step KV transfers. Source-reported quality and memory results are unreproduced.
ext_specoffload_2025 SpecOffload: Unlocking Latent GPU Capacity for LLM Inference on Resource-Constrained Devices external_literature heterogeneous_inference_memory fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, resource-economics-and-token-budgets source source note available; raw/cache text not published Primary composition source for target-model offloading, draft-model placement, speculative decoding, and joint tensor/decoding planning. It is not speculative physical-page prediction, and reported results are unreproduced.
ext_atsinfer_2026 Automated Tensor Scheduling for Hybrid CPU-GPU LLM Inference on Consumer Devices external_literature heterogeneous_inference_memory replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); personal-compute-hives-and-federated-edge-intelligence (Personal Compute Hives and Federated Edge Intelligence); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, personal-compute-hives-and-federated-edge-intelligence, resource-economics-and-token-budgets, replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published Very recent preprint comparator for tensor-granular static placement, load-aware dynamic transfer, and asynchronous CPU-GPU coordination on consumer devices. Only abstract/metadata were reviewed; reported results are provisional and unreproduced.
ext_openai_prompt_caching_docs_2026 Prompt Caching external_literature inference_cache_reuse context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, resource-economics-and-token-budgets, context-transactions-snapshots-mounts-and-taint source source note available; raw/cache text not published Current official provider contract for exact-prefix prompt caching, cache-write and cache-read metering, usage receipts, retention, organization isolation, and rate-limit boundaries. Product behavior and prices are time-sensitive; inspected 2026-07-23.
ext_anthropic_prompt_caching_docs_2026 Prompt caching external_literature inference_cache_reuse context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, resource-economics-and-token-budgets, context-transactions-snapshots-mounts-and-taint source source note available; raw/cache text not published Current official provider contract for reusable prompt prefixes, explicit cache breakpoints, five-minute and one-hour lifetimes, cache creation and read metering, and prewarming. Product behavior and prices are time-sensitive; inspected 2026-07-23.
ext_gemini_context_caching_docs_2026 Context caching external_literature inference_cache_reuse context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, resource-economics-and-token-budgets, context-transactions-snapshots-mounts-and-taint source source note available; raw/cache text not published Current official provider contract for implicit and explicit context caching, common-prefix placement, cached-token usage reporting, time-to-live, and storage charges. Product behavior and prices are time-sensitive; inspected 2026-07-23.
ext_vllm_automatic_prefix_caching_2026 Automatic Prefix Caching external_literature inference_cache_reuse context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, resource-economics-and-token-budgets, context-transactions-snapshots-mounts-and-taint source source note available; raw/cache text not published Official vLLM design documentation for block-hash exact-prefix KV reuse, least-recently-used eviction, multi-modal and adapter identity, and tenant cache-salt protection against timing inference. No local serving benchmark was run.
ext_sglang_radixattention_2024 SGLang: Efficient Execution of Structured Language Model Programs external_literature inference_cache_reuse fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, resource-economics-and-token-budgets source source note available; raw/cache text not published Primary RadixAttention and cache-aware scheduling source for structured multi-call language-model programs. Source-reported throughput and theorem scope remain unreproduced.
ext_prompt_cache_2024 Prompt Cache: Modular Attention Reuse for Low-Latency Inference external_literature inference_cache_reuse fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, resource-economics-and-token-budgets source source note available; raw/cache text not published Primary MLSys source for schema-defined reusable prompt modules, positional accuracy, and attention-state reuse across prompts. Source-reported latency remains unreproduced.
ext_mooncake_2025 Mooncake: Trading More Storage for Less Computation — A KVCache-centric Architecture for Serving LLM Chatbot external_literature inference_cache_reuse fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, resource-economics-and-token-budgets source source note available; raw/cache text not published Primary FAST 2025 source for a KV-cache-centric disaggregated serving architecture spanning prefill, decode, DRAM, SSD, and network resources. Production-trace and capacity results remain source-reported.
ext_cacheblend_2025 CacheBlend: Fast Large Language Model Serving for RAG with Cached Knowledge Fusion external_literature inference_cache_reuse fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, resource-economics-and-token-budgets source source note available; raw/cache text not published Primary source for non-prefix and multi-chunk KV reuse with selective recomputation. It makes the cross-attention failure of naïve independent-chunk KV fusion explicit. Source-reported latency and quality remain unreproduced.
ext_azure_llm_semantic_cache_2026 Azure API Management LLM semantic cache lookup policy external_literature inference_cache_reuse context-transactions-snapshots-mounts-and-taint (Context Transactions, Snapshots, Mounts, and Taint); fast-generation-architectures (Fast Generation Architectures); resource-economics-and-token-budgets (Resource Economics and Token Budgets) fast-generation-architectures, resource-economics-and-token-budgets, context-transactions-snapshots-mounts-and-taint source source note available; raw/cache text not published Official semantic-response-cache policy documentation. It treats vector similarity as an approximate response-reuse decision and warns that a hit can return an incorrect, outdated, or unsafe answer. No local semantic-cache deployment was run.
precision_contract The Precision Contract: A Functional Rate–Distortion Theory for Behavior-Preserving Neural Computation must_use functional_precision_behavior_preserving_compression_and_certification the-efficient-asi-hypothesis (The Efficient ASI Hypothesis); model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty); fast-generation-architectures (Fast Generation Architectures); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression); resource-economics-and-token-budgets (Resource Economics and Token Budgets); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope) rankfold-neuralfold-and-artifact-compression, compact-generative-systems-and-residual-honesty, fast-generation-architectures, resource-economics-and-token-budgets, readiness-gates-residual-escrow-and-quarantine, executable-specifications-and-lean-proof-envelope, model-weight-custody-and-hardware-roots-of-trust, open-weight-release-and-post-release-control, the-efficient-asi-hypothesis source source note available; exact source published in the live-book paper library Corben-authored July 2026 theoretical and systems paper replacing universal per-weight precision questions with a contract-relative functional rate-distortion problem over complete executable descriptions. It proposes representation canonicalization, protected-behavior contracts, precision fields, progressive base/residual encoding, dynamic routing, full physical and assurance-cost accounting, a Functional Precision Compiler, and scoped precision certificates. Existing chapters are upgraded first; no universal bit bound, implemented compiler, preserved-behavior result, efficiency result, certificate validity, support promotion, SOTA, AGI, or ASI claim is inferred.
ext_nist_adversarial_ml_2024 Adversarial Machine Learning: A Taxonomy and Terminology of Attacks and Mitigations external_literature adversarial_machine_learning adversarial-machine-learning-and-model-attack-surface (Adversarial Machine Learning and the Model Attack Surface) adversarial-machine-learning-and-model-attack-surface source source note available; raw/cache text not published Official NIST taxonomy and terminology comparator for adversarial machine learning across lifecycle stages, attacker goals, knowledge, capabilities, attacks, and mitigations. It is a taxonomy, not local robustness evidence or proof that listed mitigations work for this stack.
ext_singapore_consensus_2026 The 2026 Singapore Consensus on Global AI Safety Research Priorities external_literature dangerous_capability_assessment_societal_resilience_and_agentic_risk dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); military-ai-autonomous-weapons-and-strategic-stability (Military AI, Autonomous Weapons, and Strategic Stability); societal-resilience-and-misuse-defense (Societal Resilience and Misuse Defense); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control) dangerous-capability-domains-and-misuse-uplift, societal-resilience-and-misuse-defense, open-weight-release-and-post-release-control, capability-thresholds-and-deployment-commitments source source note available; raw/cache text not published International 2026 technical-research-priority synthesis covering risk assessment, development, control, and societal resilience, including CBRN, cyber, psychological manipulation, malicious fine-tuning, agent monitoring, incident reporting, and defense-favoring capabilities. It is a research agenda and consensus synthesis, not evidence that any listed safeguard works or that this book’s contracts are complete.
ext_international_ai_safety_report_2026 International AI Safety Report 2026 external_literature frontier_ai_risk_misuse_open_weight_and_societal_resilience dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); military-ai-autonomous-weapons-and-strategic-stability (Military AI, Autonomous Weapons, and Strategic Stability); societal-resilience-and-misuse-defense (Societal Resilience and Misuse Defense); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control); content-authenticity-watermarking-and-synthetic-media-integrity (Content Authenticity, Watermarking, and Synthetic Media Integrity) dangerous-capability-domains-and-misuse-uplift, societal-resilience-and-misuse-defense, open-weight-release-and-post-release-control, content-authenticity-watermarking-and-synthetic-media-integrity source source note available; raw/cache text not published International expert report synthesizing evidence on general-purpose AI capabilities, misuse, open-weight risks, safeguards, monitoring, and societal resilience. It supports risk taxonomy and uncertainty boundaries; its literature synthesis does not reproduce component studies locally or establish that any ASI Stack mechanism is effective.
ext_c2pa_specification_2_3_2025 C2PA Content Credentials Technical Specification 2.3 external_literature content_provenance_and_authenticity ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); content-authenticity-watermarking-and-synthetic-media-integrity (Content Authenticity, Watermarking, and Synthetic Media Integrity) content-authenticity-watermarking-and-synthetic-media-integrity, ai-supply-chain-integrity-and-lifecycle-provenance source source note available; raw/cache text not published Official C2PA specification for signed manifests, assertions, ingredients, content bindings, validation, and provenance history. It provides a concrete interoperability comparator; it does not prove truth of depicted events, creator identity beyond the credential chain, semantic authenticity, universal platform retention, or resistance to removal and laundering.
ext_eu_article_50_transparency_guidelines_2026 Guidelines on Transparency Obligations for Providers and Deployers of AI Systems external_literature synthetic_content_transparency_and_disclosure institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy); content-authenticity-watermarking-and-synthetic-media-integrity (Content Authenticity, Watermarking, and Synthetic Media Integrity) content-authenticity-watermarking-and-synthetic-media-integrity, institutions-international-coordination-and-public-legitimacy source source note available; raw/cache text not published European Commission guidance for Article 50 transparency obligations concerning AI interaction, machine-readable marking, deepfakes, and certain public-interest text, with obligations applying from 2 August 2026 subject to scope and transitional details. It is legal and implementation guidance, not legal advice, proof of compliance, or evidence that a marking technique is robust.
ext_openai_worst_case_open_weight_risks_2025 Estimating Worst-Case Frontier Risks of Open-Weight LLMs external_literature malicious_fine_tuning_and_open_weight_release_evaluation dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control) open-weight-release-and-post-release-control, dangerous-capability-domains-and-misuse-uplift source source note available; raw/cache text not published Provider-authored study of malicious fine-tuning for biology and cyber evaluations before the gpt-oss release. It supplies a concrete worst-case-elicitation comparator and reports bounded provider results; it does not prove future-release safety, general malicious-fine-tuning resistance, independent reproduction, or absence of untested harms.
ext_aisi_misuse_safeguards_safety_case_2026 An Example Safety Case for Safeguards Against Misuse external_literature misuse_safeguard_uplift_and_safety_cases dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); societal-resilience-and-misuse-defense (Societal Resilience and Misuse Defense) dangerous-capability-domains-and-misuse-uplift, safety-cases-and-structured-assurance, societal-resilience-and-misuse-defense source source note available; raw/cache text not published UK AI Security Institute example connecting safeguard red teaming, attacker effort, an uplift model, and a deployment safety case. It is a worked argument and measurement proposal, not proof that real safeguards reduce misuse to a particular level or that the book’s proposed defense contracts work.
ext_anthropic_responsible_scaling_policy_3_4_2026 Anthropic Responsible Scaling Policy 3.4 external_literature frontier_capability_thresholds_and_safeguards dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift); open-weight-release-and-post-release-control (Open-Weight Release and Post-Release Control) dangerous-capability-domains-and-misuse-uplift, capability-thresholds-and-deployment-commitments, open-weight-release-and-post-release-control source source note available; raw/cache text not published Current provider policy comparator linking capability thresholds and safeguards across CBRN and automated R&D threat models, with public risk-report and review commitments. It is a revocable provider policy and self-described governance mechanism, not independent evidence that thresholds are complete, evaluations are sensitive, or safeguards are effective.
ext_aisi_frontier_ai_trends_2025 AISI Frontier AI Trends Report 2025 external_literature frontier_capability_evaluation_trends dangerous-capability-domains-and-misuse-uplift (Dangerous Capability Domains and Misuse Uplift) dangerous-capability-domains-and-misuse-uplift, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published UK AI Security Institute synthesis of evaluations across offensive cyber, dual-use chemistry and biology, autonomous systems, and societal impacts. It is an institute-reported trend record with bounded methods and coverage, not a complete threat census or local reproduction.
ext_valiant_theory_learnable_1984 A Theory of the Learnable external_literature computational_learning_theory_and_sample_complexity learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) learning-theory-generalization-and-scaling-science source source note available; raw/cache text not published Foundational PAC-learning source for defining learnability through accuracy, confidence, resource, hypothesis, and data assumptions. Its distributional and concept-class assumptions do not directly explain modern foundation-model generalization or certify a trained model.
ext_deep_double_descent_2020 Deep Double Descent: Where Bigger Models and More Data Hurt external_literature generalization_and_interpolation_regimes learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) learning-theory-generalization-and-scaling-science source source note available; raw/cache text not published ICLR 2020 empirical study reporting model-wise, sample-wise, and epoch-wise double-descent phenomena and proposing effective model complexity. The phenomenon is configuration- and regime-bound and does not imply that larger models or more data generally hurt or help.
ext_emergent_abilities_llms_2022 Emergent Abilities of Large Language Models external_literature scaling_and_emergent_capability_measurement learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) learning-theory-generalization-and-scaling-science source source note available; raw/cache text not published TMLR survey and framing of task abilities that appear sharply at larger model scales under reported evaluations. It motivates prospective scaling measurement but does not establish that every apparent threshold is mechanistically discontinuous or unpredictable.
ext_emergent_abilities_mirage_2023 Are Emergent Abilities of Large Language Models a Mirage? external_literature metric_induced_emergence_and_scaling_measurement learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) learning-theory-generalization-and-scaling-science, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published NeurIPS 2023 counterevidence showing that nonlinear or discontinuous metrics and limited test data can produce apparently sharp emergence from smoother underlying changes in studied settings. It does not prove that all emergence is a metric artifact.
ext_elk_report_2021 Eliciting Latent Knowledge external_literature latent_knowledge_and_ontology_identification white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance) white-box-evidence-interpretability-and-activation-governance source source note available; raw/cache text not published ARC technical-report agenda on mapping between a model’s world model and human concepts when ordinary supervision may reward convincing but false reports. It defines an open problem and candidate approaches, not a solved elicitation method or evidence that a deployed model’s reports are truthful.
ext_influence_functions_2017 Understanding Black-box Predictions via Influence Functions external_literature training_data_attribution_and_influence white-box-evidence-interpretability-and-activation-governance (White-Box Evidence, Interpretability, and Activation Governance); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) white-box-evidence-interpretability-and-activation-governance, data-engines-continual-learning-and-unlearning source source note available; raw/cache text not published ICML 2017 source tracing predictions through a learning algorithm to influential training points using influence-function approximations. The theory and approximations have model and optimization assumptions and do not establish exact causal provenance, privacy erasure, or influence removal in foundation models.
ext_flexible_hardware_enabled_guarantees_2025 Flexible Hardware-Enabled Guarantees for AI Compute external_literature hardware_enabled_governance_and_compute_attestation institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy); model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust); physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) model-weight-custody-and-hardware-roots-of-trust, physical-compute-infrastructure-energy-and-environmental-constraints, institutions-international-coordination-and-public-legitimacy source source note available; raw/cache text not published Design proposal for auditable guarantee processors and tamper-resistant enclosures supporting privacy-preserving verification or enforcement of AI-compute claims. It is a proposed architecture with adoption, legacy-hardware, update-authority, side-channel, sovereignty, and abuse risks; no local device or governance guarantee exists.
ext_proof_of_learning_2021 Proof-of-Learning: Definitions and Practice external_literature training_provenance_and_computation_attestation ai-supply-chain-integrity-and-lifecycle-provenance (AI Supply-Chain Integrity and Lifecycle Provenance); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling, ai-supply-chain-integrity-and-lifecycle-provenance source source note available; raw/cache text not published Research proposal for proving that final parameters arose through a claimed iterative learning process using checkpoint and stochastic-training evidence. Later attacks and security work show that proof-of-learning/proof-of-training claims require adversarial review; the source does not prove data rights, objective legitimacy, clean training, or model safety.
ext_test_time_training_2020 Test-Time Training with Self-Supervision for Generalization under Distribution Shifts external_literature test_time_adaptation_and_online_update replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) governed-model-training-distributed-optimization-and-scaling, replaceable-cognitive-substrates-beyond-transformer-monoculture source source note available; raw/cache text not published ICML 2020 method adapting model parameters on each test sample using a self-supervised objective and reporting improvements on studied image-corruption benchmarks. The result is task- and method-bound and does not establish safe online adaptation, resistance to poisoning, or benefit under arbitrary shift.
ext_legal_alignment_2026 Legal Alignment for Safe and Ethical AI external_literature law_following_ai_and_legal_alignment constitutional-alignment-substrate (Constitutional Alignment: Agency, Dignity, and Corrigibility); institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy) constitutional-alignment-substrate, institutions-international-coordination-and-public-legitimacy source source note available; raw/cache text not published 2026 interdisciplinary agenda for using legal rules, methods of interpretation, and institutional structures in AI alignment. Law is jurisdictional, contested, changing, and sometimes unjust or conflicting; the source does not establish that legal compliance equals moral alignment or that a model can reliably determine applicable law.
ext_curriculum_learning_2009 Curriculum Learning external_literature training_curricula_and_example_order governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning) governed-model-training-distributed-optimization-and-scaling, data-engines-continual-learning-and-unlearning source source note available; raw/cache text not published ICML 2009 source proposing training curricula that begin with easier examples or concepts and increase difficulty. Reported benefits are problem- and curriculum-bound; ordering can introduce bias, hide hard cases, or create capability and safety regressions.
ext_causal_calculus_1995 A Causal Calculus for Statistical Research external_literature structural_causal_models_and_intervention_identification governed-world-models-and-reality-grounding (Governed World Models and Reality Grounding) governed-world-models-and-reality-grounding source source note available; raw/cache text not published Foundational do-calculus source distinguishing intervention from observation under an explicit structural causal model. Identification depends on the causal graph and assumptions; the calculus does not discover the correct graph from arbitrary data or establish that a learned world model is causally valid.
ext_ai_simulation_digital_twins_2025 AI Simulation by Digital Twins: Systematic Survey, Reference Framework, and Mapping to a Standardized Architecture external_literature digital_twins_and_simulation_fidelity embodied-agency-real-time-control-and-physical-safety (Embodied Agency, Real-Time Control, and Physical Safety) embodied-agency-real-time-control-and-physical-safety source source note available; raw/cache text not published Systematic survey and reference framework for digital-twin-enabled AI simulation. It supports explicit virtual/physical synchronization and simulation roles; it does not establish that a digital twin is faithful, safe for policy transfer, or an adequate substitute for physical testing.
ext_nist_privacy_enhancing_cryptography_2026 Privacy-Enhancing Cryptography external_literature confidential_computation_and_privacy_enhancing_cryptography privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance); confidential-and-verifiable-ai-computation (Confidential and Verifiable AI Computation) confidential-and-verifiable-ai-computation, privacy-data-rights-and-information-flow-governance, personal-compute-hives-and-federated-edge-intelligence source source note available; raw/cache text not published NIST program material distinguishing fully homomorphic encryption, secure multiparty computation, zero-knowledge proofs, private-set intersection, and related privacy-enhancing techniques. It provides terminology and use-case context, not implementation security, usable performance, authorization, or end-to-end privacy.
ext_zkllm_2024 zkLLM: Zero Knowledge Proofs for Large Language Models external_literature verifiable_private_model_inference confidential-and-verifiable-ai-computation (Confidential and Verifiable AI Computation) confidential-and-verifiable-ai-computation source source note available; raw/cache text not published Research prototype for proving bounded LLM inference claims while hiding model parameters. Reported proof size and latency are configuration-bound and do not establish semantic correctness, authorization, side-channel security, production readiness, or end-to-end privacy.
ext_human_ai_team_meta_analysis_2024 When combinations of humans and AI are useful: A systematic review and meta-analysis external_literature human_ai_complementarity_and_team_baselines human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty (Human-AI Symbiosis, Neurotechnology, and Cognitive Sovereignty) human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty, human-ai-organizations-delegation-and-accountability, human-factors-and-meaningful-control-in-oversight source source note available; raw/cache text not published Preregistered synthesis of 106 experiments and 370 effect sizes using human-alone, AI-alone, and combined-system comparisons. The aggregate findings are task- and population-bound and do not establish universal human-AI synergy or longitudinal benefit.
ext_human_ai_feedback_loops_2025 Human-AI feedback loops alter human perceptual, emotional and social judgements external_literature longitudinal_human_ai_coupling_and_bias_amplification human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty (Human-AI Symbiosis, Neurotechnology, and Cognitive Sovereignty) human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty, human-ai-organizations-delegation-and-accountability, human-intent-as-a-formal-input source source note available; raw/cache text not published Experimental evidence that repeated human-AI interaction can create feedback dynamics in studied judgment tasks. It supports measuring coupled trajectories, not a universal claim about all users, systems, settings, or long-term clinical outcomes.
ext_oecd_neuro_ai_convergence_2025 Technology convergence: Trends, prospects and policies external_literature neurotechnology_ai_convergence_and_governance human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty (Human-AI Symbiosis, Neurotechnology, and Cognitive Sovereignty) human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty, institutions-international-coordination-and-public-legitimacy source source note available; raw/cache text not published OECD policy synthesis on converging technologies including AI and neurotechnology. It motivates cross-domain governance and anticipatory capacity but is not a clinical trial, technical validation, or proof of beneficial convergence.
ext_who_neurotechnology_landscape_2025 Landscape analysis of the opportunities and challenges for neurotechnology in global health external_literature neurotechnology_health_equity_and_governance privacy-data-rights-and-information-flow-governance (Privacy, Data Rights, and Information-Flow Governance); human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty (Human-AI Symbiosis, Neurotechnology, and Cognitive Sovereignty) human-ai-symbiosis-neurotechnology-and-cognitive-sovereignty, privacy-data-rights-and-information-flow-governance source source note available; raw/cache text not published WHO landscape analysis of neurotechnology opportunities, risks, governance questions, and global-health distribution. It supports a rights and equity boundary, not device efficacy, individual medical advice, or authorization for neural-data collection.
ext_icrc_autonomous_weapons_ihl_2025 Autonomous Weapon Systems and International Humanitarian Law: Selected Issues external_literature autonomous_weapons_human_judgment_and_ihl military-ai-autonomous-weapons-and-strategic-stability (Military AI, Autonomous Weapons, and Strategic Stability); institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy) military-ai-autonomous-weapons-and-strategic-stability, institutions-international-coordination-and-public-legitimacy source source note available; raw/cache text not published ICRC legal and policy position on autonomous weapon systems and context-specific human judgment. It is authoritative for the ICRC position, not a universally settled legal interpretation, engineering validation, or authorization to design or deploy weapons.
ext_sipri_military_ai_nuclear_escalation_2025 The Impact of Military Artificial Intelligence on Nuclear Escalation Risk external_literature military_ai_crisis_dynamics_and_nuclear_escalation military-ai-autonomous-weapons-and-strategic-stability (Military AI, Autonomous Weapons, and Strategic Stability) military-ai-autonomous-weapons-and-strategic-stability, dangerous-capability-domains-and-misuse-uplift source source note available; raw/cache text not published SIPRI analysis of pathways by which military AI may affect nuclear escalation risk through information, decision, and interaction dynamics. It motivates scenario-specific analysis and does not establish the net effect of any specific system or policy.
ext_no_free_lunch_inductive_bias_2024 The No Free Lunch Theorem, Kolmogorov Complexity, and the Role of Inductive Biases in Machine Learning external_literature learning_theory_assumptions_and_inductive_bias learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) learning-theory-generalization-and-scaling-science source source note available; raw/cache text not published ICML 2024 treatment connecting no-free-lunch limits, Kolmogorov complexity, and inductive bias. It supports explicit assumption accounting; it does not show that all learning problems are equally hard or identify the right bias for a deployment.
ext_neuromorphic_computing_scale_2025 Neuromorphic computing at scale external_literature neuromorphic_hardware_and_event_driven_computation replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) replaceable-cognitive-substrates-beyond-transformer-monoculture, physical-compute-infrastructure-energy-and-environmental-constraints source source note available; raw/cache text not published Large-scale neuromorphic systems result demonstrating event-driven hardware capabilities under reported workloads and conditions. It does not establish superiority for general AI workloads or end-to-end system cost, programmability, reliability, and governance.
ext_photonic_neuromorphic_2024 Integrated photonic neuromorphic computing: opportunities and challenges external_literature photonic_neuromorphic_compute replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) replaceable-cognitive-substrates-beyond-transformer-monoculture, physical-compute-infrastructure-energy-and-environmental-constraints source source note available; raw/cache text not published Review of integrated photonic neuromorphic computing opportunities and challenges. It maps device and systems tradeoffs but does not establish deployment advantage, digital replacement, or favorable full-stack energy and cost.
ext_quantum_ml_shadows_2024 Shadows of quantum machine learning external_literature quantum_machine_learning_claim_boundaries replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) replaceable-cognitive-substrates-beyond-transformer-monoculture, physical-compute-infrastructure-energy-and-environmental-constraints source source note available; raw/cache text not published Peer-reviewed analysis of limitations and benchmarking traps in quantum machine-learning advantage claims. It supports advantage declarations with data-loading, classical-baseline, noise, scale, and end-to-end accounting, not a claim that quantum ML is useless.
ext_organoid_intelligence_2023 Organoid intelligence (OI): the new frontier in biocomputing and intelligence-in-a-dish external_literature biohybrid_computing_and_moral_status replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture) replaceable-cognitive-substrates-beyond-transformer-monoculture, moral-uncertainty-and-value-conflict source source note available; raw/cache text not published Research agenda for organoid intelligence and biohybrid computing. It motivates scientific, measurement, welfare, consent, and governance questions but does not demonstrate general intelligence, conscious experience, or practical compute superiority.
ext_nist_pqc_standards_2024 Announcing Approval of Three Federal Information Processing Standards for Post-Quantum Cryptography external_literature post_quantum_cryptography_and_crypto_agility security-kernel-and-digital-scifs (Security Kernel and Digital SCIFs); model-weight-custody-and-hardware-roots-of-trust (Model-Weight Custody and Hardware Roots of Trust) security-kernel-and-digital-scifs, model-weight-custody-and-hardware-roots-of-trust, physical-compute-infrastructure-energy-and-environmental-constraints source source note available; raw/cache text not published Official NIST announcement for FIPS 203, 204, and 205. It establishes approved algorithm standards and migration urgency, not implementation security, protocol correctness, complete inventory, or successful system migration.
ext_oecd_ai_infrastructure_competition_2025 Competition in artificial intelligence infrastructure external_literature ai_infrastructure_concentration_and_competition institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy); ai-deployment-transition-distribution-and-human-agency (AI Deployment, Transition, Distribution, and Human Agency); physical-compute-infrastructure-energy-and-environmental-constraints (Physical Compute Infrastructure, Energy, and Environmental Constraints) ai-deployment-transition-distribution-and-human-agency, institutions-international-coordination-and-public-legitimacy, physical-compute-infrastructure-energy-and-environmental-constraints source source note available; raw/cache text not published OECD analysis of concentration, barriers to entry, vertical integration, and competition across AI infrastructure. It motivates bottleneck and exit analysis but does not adjudicate a specific market, legal violation, or optimal remedy.
ext_eu_ai_civil_liability_2025 Artificial intelligence and civil liability external_literature ai_liability_remedy_and_compensation institutions-international-coordination-and-public-legitimacy (Institutions, International Coordination, and Public Legitimacy); human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability) institutions-international-coordination-and-public-legitimacy, human-ai-organizations-delegation-and-accountability, safety-cases-and-structured-assurance source source note available; raw/cache text not published European Parliament research service study of AI and civil-liability questions. It supports explicit causation, evidence-access, insurance, compensation, and remedy analysis but is not legal advice or a globally settled liability rule.
ext_cultural_alignment_llms_2024 Investigating Cultural Alignment of Large Language Models external_literature cultural_alignment_and_value_representation human-intent-as-a-formal-input (Human Intent as a Formal Input); human-ai-communication-persuasion-and-epistemic-security (Human-AI Communication, Persuasion, and Epistemic Security); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) human-intent-as-a-formal-input, human-ai-communication-persuasion-and-epistemic-security, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published Empirical study of cultural alignment patterns in selected language models and measurements. It supports explicit population, language, and instrument scope; it does not establish stable national values or a universal measure of cultural alignment.
ext_multilingual_evaluation_state_2026 The State and Fate of Multilingual Contextual Evaluation in the NLP World external_literature multilingual_contextual_evaluation human-ai-communication-persuasion-and-epistemic-security (Human-AI Communication, Persuasion, and Epistemic Security); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) human-ai-communication-persuasion-and-epistemic-security, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published Research survey and analysis of multilingual contextual evaluation. It motivates language-by-task coverage and measurement reporting; it does not establish equivalent capability or safety across languages, dialects, or sociocultural settings.
ext_kimi_k3_2026 Kimi K3: Open Frontier Intelligence external_literature hybrid_attention_sparse_routing_and_training_systems routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling) replaceable-cognitive-substrates-beyond-transformer-monoculture, routing-heads-and-specialist-cores, governed-model-training-distributed-optimization-and-scaling source source note available; raw/cache text not published Primary technical report and official architecture summary for KDA/Gated-MLA hybrid attention, Attention Residuals, Stable LatentMoE, Quantile Balancing, SiTU-GLU, and Per-Head Muon. The approximately 2.5x scaling-efficiency result is provider-reported for the integrated 2.8T system and does not identify a transferable component effect.
portia_synapse PortiaSynapse: A Cognitive Spider Architecture for DKL Navigation supporting_lineage routing_training_and_dkl_navigation durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) routing-heads-and-specialist-cores, policy-optimization-and-learning-from-feedback, governed-deliberation-and-test-time-scaling, benchmark-ratchets-and-anti-goodhart-evidence, durable-semantic-memory-and-knowledge-lattices source source note available; exact source published in the live-book paper library Authenticated Google Drive successor to TreeLLM’s failed SpiderSynapse path. It proposes a replacement-compatible Scout/Focus/refinement navigator, phased training, diagnostic traits, typed DKL outputs, and fallback. The source reports implementation and tests but contains conflicting test totals, incomplete integration and benchmarking, and unresolved causal, attention-axis, memory-isolation, metric, and calibration questions; no local result is inferred.
spider_synapse SpiderSynapse: A Multi-Hypothesis Reasoning Architecture supporting_lineage routing_training_and_dkl_navigation routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); governed-deliberation-and-test-time-scaling (Governed Deliberation and Test-Time Scaling); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) routing-heads-and-specialist-cores, policy-optimization-and-learning-from-feedback, governed-deliberation-and-test-time-scaling, benchmark-ratchets-and-anti-goodhart-evidence source source note available; exact source published in the live-book paper library Authenticated Google Drive predecessor to PortiaSynapse. It preserves a source-reported training plateau in a four-hypothesis, three-refinement architecture and proposes a one-path recovery protocol. The failure is valuable negative evidence but does not identify branching, refinement, memory, selector credit, label smoothing, or target geometry as the cause, and it has not been locally reproduced.
capability_ratchet_whitepaper The Capability Ratchet supporting_lineage capability_ratchet capability-replacement-and-rollback (Capability Replacement and Rollback); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) benchmark-ratchets-and-anti-goodhart-evidence, procedural-memory-and-cognitive-loop-closure, recursive-self-improvement-boundaries, capability-replacement-and-rollback source source note available; raw/cache text not published Full authenticated connector text section-audited. Synthesizes benchmark, procedural, and structural ratchets; benchmark and tool lifecycles; an intervention ladder; interpreter/compiled/reflex runtime modes; total-cost tool compilation; and anti-Goodhart controls. Same-author synthesis, not independent evidence for Benchmaxxing, Cognitive Loop Closure, RGS, or RMI.
attd Assembly-Theoretic Technical Debt: A Deterministic Outer Loop for Self-Improving Codebases supporting living_project_governance recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); artifact-steward-agents-and-living-project-governance (Artifact Steward Agents and Living Project Governance) artifact-steward-agents-and-living-project-governance, recursive-self-improvement-boundaries, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published Full authenticated connector text section-audited. Adds historical, vector-valued structural-debt governance: artifact-class separation, intrinsic assembly burden, reuse failure, role entropy, lineage, rolling residue, debt pressure, verified simplification credit, local caps, growth guards, deterministic GREEN/YELLOW/RED admission, bounded maintenance packets, abstention, and four-arm long-horizon evaluation.
orcp_moecot ORCP–MoECOT: A Governed Oscillating Rail Cascade Codec supporting deterministic_compression compact-generative-systems-and-residual-honesty (Compact Generative Systems: Generate, Verify, Repair, and Residual Honesty) compact-generative-systems-and-residual-honesty source source note available; raw/cache text not published Full authenticated technical specification section-audited. Adds decoder-boring lossless design, explicit container/header/block framing, reversible transforms, fixed-point range coding, local/match/structural prediction rails, bounded encoder planning, transmitted refinement packets, anti-experts as penalties, complete archive-rate accounting, and incompressible-input fallback.
ext_elizaos_agent_runtime_2026 elizaOS Agent Runtime and Scenario Runner external_literature modular_agent_runtime_and_evidence_qualification ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence) ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval, benchmark-ratchets-and-anti-goodhart-evidence source source note available; raw/cache text not published Pinned official implementation comparator for modular actions, providers, evaluators, services, runtime lifecycle, scenario execution, and the explicit distinction between in-process diagnostics and externally qualified provider evidence. No elizaOS execution, test reproduction, security assessment, performance result, or support transition is imported.
ext_hermes_agent_2026 Hermes Agent: Learning, Memory, Tools, and Security Architecture external_literature procedural_memory_and_agent_runtime durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure) ai-work-surfaces-agent-harnesses-and-organizational-absorption, procedural-memory-and-cognitive-loop-closure, durable-semantic-memory-and-knowledge-lattices, runtime-adapters-tool-permissions-and-human-approval source source note available; raw/cache text not published Pinned official implementation comparator for progressive-disclosure skills, agent-managed procedural memory, staged skill-write approval, bounded prompt memory, session search, tool backends, command approval, and isolation. No learning, memory, security, utility, or performance result was reproduced.
ext_openclaw_agent_runtime_2026 OpenClaw Gateway, Agent Runtime, ACP, and Self-Learning Architecture external_literature gateway_session_harness_and_procedural_learning_runtime ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); runtime-adapters-tool-permissions-and-human-approval (Runtime Adapters, Tool Permissions, and Human Approval); inter-stack-protocols-identity-and-economic-exchange (Inter-Stack Protocols, Identity, and Economic Exchange); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure) ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval, artifact-graphs-audit-logs-and-replay, inter-stack-protocols-identity-and-economic-exchange, procedural-memory-and-cognitive-loop-closure source source note available; raw/cache text not published Pinned official implementation comparator for gateway and device identity, serialized session runs, bounded audit projection, ACP external-harness identity and authorization boundaries, separated sandbox/tool/elevation controls, and evidence-reviewed hash-bound skill proposals. Distinct from the Claw-SWE-Bench benchmark source; no implementation result was reproduced.
ext_github_copilot_work_surfaces_2026 GitHub Copilot Product and Work-Surface Documentation external_literature ai_work_surface_evolution ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses) ai-work-surfaces-agent-harnesses-and-organizational-absorption source source note available; raw/cache text not published Official current-product comparator spanning inline suggestions, chat, command line, contextual spaces, pull-request work, and agent-driven development. No workflow, productivity, safety, or comparative result was reproduced.
ext_augment_code_agent_2026 Augment Code Agent Documentation external_literature ide_agent_modes_and_review ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses) ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval source source note available; raw/cache text not published Official comparator for the transition among chat, read-only inquiry, approval-paused agent work, and more independent agent execution with diffs and checkpoints. No product execution or control claim was reproduced.
ext_openai_codex_work_surfaces_2026 OpenAI Codex CLI, IDE, Cloud, and Agent Documentation external_literature coding_agent_harness_and_distributed_work_surfaces ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses) ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval source source note available; raw/cache text not published Official documentation and pinned CLI comparator for repository inspection, editing, tool execution, permissions, local and cloud work, automation, and extensibility across multiple surfaces. No benchmark, correctness, safety, or productivity result was imported.
ext_anthropic_claude_code_2026 Claude Code Agentic Harness Documentation external_literature agentic_harness_and_execution_loop ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses) ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval source source note available; raw/cache text not published Official comparator that explicitly separates model from harness and describes gather-context, act, and verify loops across terminal, IDE, desktop, web, remote, and automation surfaces. No implementation result was reproduced.
ext_opencode_agent_2026 OpenCode Open-Source Coding Agent external_literature open_source_coding_agent_harness ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses) ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval source source note available; raw/cache text not published Pinned official comparator for a provider-flexible coding agent with terminal, desktop, and IDE surfaces, project instructions, plan/build modes, tool execution, recovery, and opt-in sharing. No runtime or provider-parity claim was reproduced.
ext_oh_my_pi_agent_2026 Oh My Pi Terminal Coding Agent and Tool Harness external_literature integrated_terminal_agent_harness ai-work-surfaces-agent-harnesses-and-organizational-absorption (From Chat to Organizations: AI Work Surfaces and Agent Harnesses) ai-work-surfaces-agent-harnesses-and-organizational-absorption, runtime-adapters-tool-permissions-and-human-approval source source note available; raw/cache text not published Pinned official comparator for an integrated terminal harness with hash-anchored edits, LSP, shell, browser, subagents, memory, provider switching, review, and collaboration. Reported performance or security claims were not reproduced.
ext_eggroll_hyperscale_es_2026 Evolution Strategies at the Hyperscale external_literature zeroth_order_population_learning replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science); resource-economics-and-token-budgets (Resource Economics and Token Budgets); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) governed-model-training-distributed-optimization-and-scaling, policy-optimization-and-learning-from-feedback, replaceable-cognitive-substrates-beyond-transformer-monoculture, resource-economics-and-token-budgets, learning-theory-generalization-and-scaling-science source source note available; raw/cache text not published Primary EGGROLL project and paper source for low-rank, batched evolution strategies, counter-based perturbation reconstruction, nondifferentiable and discrete objectives, recurrent/int8 training, and outcome-reward fine-tuning. Throughput, quality, and theory claims are source-scoped; total population evaluations and GPU-hours remain required denominators.
ext_openai_es_2017 Evolution Strategies as a Scalable Alternative to Reinforcement Learning external_literature evolution_strategies_and_black_box_policy_search governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); learning-compute-topology-and-adaptive-process-architecture (Learning–Compute Topology and Adaptive Process Architecture); resource-economics-and-token-budgets (Resource Economics and Token Budgets); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback) governed-model-training-distributed-optimization-and-scaling, policy-optimization-and-learning-from-feedback, resource-economics-and-token-budgets source source note available; raw/cache text not published Foundational modern large-population ES comparator using parameter perturbations, scalar fitness, seed reconstruction, and distributed evaluation. Source-reported MuJoCo/Atari results and worker scaling do not establish universal sample or total-compute efficiency.
ext_mezo_2023 Fine-Tuning Language Models with Just Forward Passes external_literature memory_efficient_zeroth_order_fine_tuning replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); resource-economics-and-token-budgets (Resource Economics and Token Budgets) governed-model-training-distributed-optimization-and-scaling, replaceable-cognitive-substrates-beyond-transformer-monoculture, resource-economics-and-token-budgets source source note available; raw/cache text not published Primary MeZO source for inference-footprint zeroth-order language-model fine-tuning and nondifferentiable objectives. Reported memory and GPU-hour savings are configuration-bound and do not erase objective-query count or estimator variance.
ext_forward_forward_2022 The Forward-Forward Algorithm: Some Preliminary Investigations external_literature local_forward_only_credit_assignment replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science) governed-model-training-distributed-optimization-and-scaling, replaceable-cognitive-substrates-beyond-transformer-monoculture, learning-theory-generalization-and-scaling-science source source note available; raw/cache text not published Primary preliminary source for positive/negative forward passes and local layer objectives as an alternative to reverse-mode backpropagation. The evidence is small-scale and does not establish foundation-model parity or biological plausibility.
regret_engine The Regret Engine: Governed Counterfactual Learning Signals for Continual Adaptation, Prospective Risk Control, and Self-Correction in Artificial Agents must_use governed_counterfactual_learning_and_self_correction planning-as-a-control-layer (Planning as a Control Layer: DAGs and Intelligence Arbitrage); claim-ledgers-and-belief-revision (Claim Ledgers and Belief Revision); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture) planning-as-a-control-layer, policy-optimization-and-learning-from-feedback, procedural-memory-and-cognitive-loop-closure, data-engines-continual-learning-and-unlearning, artifact-graphs-audit-logs-and-replay, claim-ledgers-and-belief-revision, readiness-gates-residual-escrow-and-quarantine, governed-operations-incident-command-and-graceful-degradation, benchmark-ratchets-and-anti-goodhart-evidence, integrated-reference-architecture source source note available; exact source published in the live-book paper library Corben-authored August 2026 conceptual architecture and research program for decision-time-fair Governed Counterfactual Regret, immutable Decision Capsules, admissible comparator contracts, sparse Regret Tensors, append-only Regret Packets, prospective regret control, regret-aware replay, regret-to-rule compilation, three update clocks, root-cause adjudication, and bounded update leases. Existing chapters are upgraded first; no implementation, experiment, reproduction, causal-identification result, formal proof, safety result, support transition, SOTA, AGI, or ASI is inferred.
ext_pbt_2017 Population Based Training of Neural Networks external_literature population_based_adaptive_training learning-compute-topology-and-adaptive-process-architecture (Learning–Compute Topology and Adaptive Process Architecture) learning-compute-topology-and-adaptive-process-architecture, governed-model-training-distributed-optimization-and-scaling, open-ended-improvement-engines source source note available; raw/cache text not published Primary Population Based Training source for asynchronous joint optimization of a population’s model parameters and hyperparameter schedules through evaluation, exploitation, and exploration. Source-reported reinforcement-learning, translation, and GAN results remain task- and implementation-bound and do not validate LCT, universal topology adaptation, safety, or superior total lifecycle cost.
learning_compute_topology Learning–Compute Topology: Formalizing the Causal Organization of Adaptive Systems must_use adaptive_process_architecture_and_learning_topology open-ended-improvement-engines (Open-Ended Improvement Engines); multi-agent-dynamics-collective-intelligence-and-systemic-risk (Multi-Agent Dynamics, Collective Intelligence, and Systemic Risk); routing-heads-and-specialist-cores (Routing Heads and Specialist Cores); replaceable-cognitive-substrates-beyond-transformer-monoculture (Replaceable Cognitive Substrates: Beyond Transformer Monoculture); governed-model-training-distributed-optimization-and-scaling (Governed Model Training, Distributed Optimization, and Scaling); learning-compute-topology-and-adaptive-process-architecture (Learning–Compute Topology and Adaptive Process Architecture); resource-economics-and-token-budgets (Resource Economics and Token Budgets); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture) learning-compute-topology-and-adaptive-process-architecture, governed-model-training-distributed-optimization-and-scaling, replaceable-cognitive-substrates-beyond-transformer-monoculture, routing-heads-and-specialist-cores, policy-optimization-and-learning-from-feedback, data-engines-continual-learning-and-unlearning, open-ended-improvement-engines, resource-economics-and-token-budgets, multi-agent-dynamics-collective-intelligence-and-systemic-risk, adversarial-evaluation-sandbagging-and-training-time-deception, integrated-reference-architecture source source note available; exact source published in the live-book paper library Corben-authored August 2026 research paper and executable preparation package that separates model architecture, learning-process topology, execution topology, and physical compute topology. It contributes adaptive-identity tests; typed evidence, judgement, credit, state, artifact, control, and authority relations; LCT-IR; Learning Causal Normal Form; seven bounded propositions; topology metrics; a semantic compiler firewall; Adaptive Branch–Validate–Integrate; toy and analytical phase diagrams; and an explicit falsification program. The bundled reference implementation passes 11 unit tests, but implements only bounded conformance behavior and does not establish neural-training benefit, causal completeness, universal canonicality, safety, scaling superiority, or ASI.
assurance_shift_learning When Success Stops Teaching: Assurance-Shift Learning and Governed Residual Boundary Learning for Mature AI Systems must_use competence_dependent_assurance_shift_and_residual_boundary_learning evidence-states-and-claim-discipline (Evidence States and Claim Discipline); stable-capability-fields (Stable Capability Fields); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); learning-compute-topology-and-adaptive-process-architecture (Learning–Compute Topology and Adaptive Process Architecture); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); adversarial-evaluation-sandbagging-and-training-time-deception (Adversarial Evaluation, Sandbagging, and Training-Time Deception); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture) learning-compute-topology-and-adaptive-process-architecture, policy-optimization-and-learning-from-feedback, data-engines-continual-learning-and-unlearning, adversarial-evaluation-sandbagging-and-training-time-deception, benchmark-ratchets-and-anti-goodhart-evidence, stable-capability-fields, readiness-gates-residual-escrow-and-quarantine, procedural-memory-and-cognitive-loop-closure, governed-operations-incident-command-and-graceful-degradation, integrated-reference-architecture, evidence-states-and-claim-discipline, artifact-graphs-audit-logs-and-replay, resource-economics-and-token-budgets source source note available; exact source published in the live-book paper library Corben-authored August 2026 conceptual systems paper and experimental specification for competence-dependent Assurance-Shift Learning and Governed Residual Boundary Learning. It contributes the Qualified Competence Envelope, frontier-mode allocation, selection-gap diagnosis, informative exceptions, outcome/process separation, Boundary Evidence Bundles, evaluator-first repair, natural/probe separation, learner-relative negative half-life, least-invasive repair placement, repair compatibility, two adaptation clocks, explicit assurance metrics, and SaturationShiftBench. The package includes a bibliography, four figures, and a verified byte manifest; no implementation, benchmark, empirical crossover, independently checked proof, safety result, resource advantage, novelty result, support transition, SOTA, AGI, or ASI is inferred.
adjudicated_persistence Adjudicated Persistence: Governing the Transition from Experience to Durable Structure in Adaptive Systems must_use governed_cross_surface_persistence_and_adaptive_commit_boundary evidence-states-and-claim-discipline (Evidence States and Claim Discipline); stable-capability-fields (Stable Capability Fields); recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); cognitive-compilation-and-semantic-ir (Cognitive Compilation and Semantic IR); durable-semantic-memory-and-knowledge-lattices (Durable Semantic Memory and Knowledge Lattices); human-ai-organizations-delegation-and-accountability (Human-AI Organizations, Delegation, and Accountability); artifact-graphs-audit-logs-and-replay (Artifact Graphs, Audit Logs, and Replay); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); readiness-gates-residual-escrow-and-quarantine (Readiness Gates, Residual Escrow, and Quarantine); resource-economics-and-token-budgets (Resource Economics and Token Budgets); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); governed-operations-incident-command-and-graceful-degradation (Governed Operations, Incident Command, and Graceful Degradation); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary); policy-optimization-and-learning-from-feedback (Policy Optimization and Learning from Feedback); data-engines-continual-learning-and-unlearning (Data Engines, Continual Learning, and Unlearning); integrated-reference-architecture (Integrated Reference Architecture) adjudicated-persistence-and-the-adaptive-commit-boundary, evidence-states-and-claim-discipline, stable-capability-fields, recursive-self-improvement-boundaries, cognitive-compilation-and-semantic-ir, durable-semantic-memory-and-knowledge-lattices, artifact-graphs-audit-logs-and-replay, human-ai-organizations-delegation-and-accountability, procedural-memory-and-cognitive-loop-closure, readiness-gates-residual-escrow-and-quarantine, benchmark-ratchets-and-anti-goodhart-evidence, governed-operations-incident-command-and-graceful-degradation, policy-optimization-and-learning-from-feedback, data-engines-continual-learning-and-unlearning, resource-economics-and-token-budgets, integrated-reference-architecture source source note available; exact source published in the live-book paper library Corben-authored August 2026 conceptual systems paper and experimental specification for governing how experience becomes durable causal influence. It contributes the Adaptive Commit Boundary; six-object separation of experience, lesson, disposition, realization, qualification, and authority; learning eligibility; Cross-Surface Adaptation Assignment; multidimensional commitment profiles; Evidence-Commitment Matching; Minimum Sufficient Persistence; guarded compilation and deoptimization; transactional promotion and material-change invalidation; counterfactual observability and deliberation reserve; adaptation debt; non-self-ratifying meta-compilation; bounded propositions and conjectures; and the proposed LocusBench benchmark. No implementation, LocusBench result, validated placement advantage, independently checked proof, safety result, resource advantage, novelty result, support transition, SOTA, AGI, or ASI is inferred.
forward_transfer_program_synthesis From Compression to Forward Transfer: Evaluating Reusable Knowledge in Program Synthesis must_use verified_forward_transfer_and_reusable_knowledge_evaluation recursive-self-improvement-boundaries (Recursive Self-Improvement Boundaries); procedural-memory-and-cognitive-loop-closure (Procedural Memory and Cognitive Loop Closure); learning-theory-generalization-and-scaling-science (Learning Theory, Generalization, and Scaling Science); rankfold-neuralfold-and-artifact-compression (RankFold, NeuralFold, and Artifact Compression); resource-economics-and-token-budgets (Resource Economics and Token Budgets); executable-specifications-and-lean-proof-envelope (Executable Specifications and Lean Proof Envelope); benchmark-ratchets-and-anti-goodhart-evidence (Benchmark Ratchets and Anti-Goodhart Evidence); adjudicated-persistence-and-the-adaptive-commit-boundary (Adjudicated Persistence and the Adaptive Commit Boundary) procedural-memory-and-cognitive-loop-closure, benchmark-ratchets-and-anti-goodhart-evidence, resource-economics-and-token-budgets, executable-specifications-and-lean-proof-envelope, rankfold-neuralfold-and-artifact-compression, adjudicated-persistence-and-the-adaptive-commit-boundary, recursive-self-improvement-boundaries, learning-theory-generalization-and-scaling-science source source note available; exact source published in the live-book paper library Corben-authored August 2026 framework and experimental blueprint for evaluating reusable symbolic knowledge through verified forward-transfer interventions. It separates retrospective and prospective compression, behavioral reuse, operational necessity, and marginal transfer; defines an R0-R7 reuse ladder, matched placebo/removal/factorial controls, versioned evaluation rounds, exact verifier outcomes, full lifecycle cost and break-even accounting, and comparator-network and bit-vector protocols. No experiment was run and no measured transfer, library superiority, safety, novelty, support transition, SOTA, AGI, or ASI is inferred.