Skip to content

Documentation

Component Stability

Stability classification for the 213 components in the registry snapshot below, last audited 2026-04-20. G6 has 269 components in total, so 56 carry no classification yet. Below the registry table: how G6 builds itself — the same self-training methodology used on external benchmarks, applied to the system’s own production readiness.

213
Total
165
Production (Tier 6)
2
Stable
26
Beta (Tier 4)
20
Experimental

This is a dated snapshot, not a live status board. The registry was last audited on 2026-04-20 and is not re-checked on page load. It classifies 213 of G6’s 269 components; the remaining 56 are unclassified, which means unassessed — not that they passed.

Component Stability Since Notes
adapt_audio PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_automl PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_bayesian PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_bias PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_blender PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_comfyui PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_creative_api PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_diagrams PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_django PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_eurisko PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_experta PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_ffmpeg PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_generative_art PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_healing PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_image PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_instructor PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_keras PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_learning PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_library_synthesizer PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_memory PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_optimisation PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_pandas PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_physical_ai PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_pygad PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_pytorch PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_rq PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_sklearn PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_synth PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_synthetic_world PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_trm PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_ui_design PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_visualisation PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_voice PRODUCTION 2026-04-20 T6: production-ready with observability
adapt_webpage PRODUCTION 2026-04-20 T6: production-ready with observability
affordance_kb PRODUCTION 2026-04-20 T6: production-ready with observability
agent_autogen PRODUCTION 2026-04-20 T6: production-ready with observability
agent_langchain PRODUCTION 2026-04-20 T6: production-ready with observability
agent_langgraph PRODUCTION 2026-04-20 T6: production-ready with observability
agent_nanoclaw PRODUCTION 2026-04-20 T6: production-ready with observability
agent_openai PRODUCTION 2026-04-20 T6: production-ready with observability
agent_openclaw PRODUCTION 2026-04-20 T6: production-ready with observability
agent_smolagents PRODUCTION 2026-04-20 T6: production-ready with observability
align_artifacts PRODUCTION 2026-04-20 T6: production-ready with observability
align_coconstructive PRODUCTION 2026-04-20 T6: production-ready with observability
align_csf PRODUCTION 2026-04-20 T6: production-ready with observability
align_evals PRODUCTION 2026-04-20 T6: production-ready with observability
align_health PRODUCTION 2026-04-20 T6: production-ready with observability
align_prompt_library PRODUCTION 2026-04-20 T6: production-ready with observability
align_specs PRODUCTION 2026-04-20 T6: production-ready with observability
align_verbsamp PRODUCTION 2026-04-20 T6: production-ready with observability
auto_component_integrator PRODUCTION 2026-04-20 T6: production-ready with observability
auto_engineer PRODUCTION 2026-04-20 T6: production-ready with observability
autonomous_orchestrator PRODUCTION 2026-04-20 T6: production-ready with observability
autonomy_governor PRODUCTION 2026-04-20 T6: production-ready with observability
cog_arch_actr PRODUCTION 2026-04-20 T6: production-ready with observability
cog_arch_aixi PRODUCTION 2026-04-20 T6: production-ready with observability
cog_arch_dgm PRODUCTION 2026-04-20 T6: production-ready with observability
cog_arch_gps PRODUCTION 2026-04-20 T6: production-ready with observability
cog_arch_soar PRODUCTION 2026-04-20 T6: production-ready with observability
component_creator PRODUCTION 2026-04-20 T6: production-ready with observability
context_engine PRODUCTION 2026-04-20 T6: production-ready with observability
csf_strategy PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_ace PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_claude_context PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_claude_mem PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_cognee PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_colbert PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_contexthub PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_fenic PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_geo PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_git PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_langextract PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_library_mapper PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_markitdown PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_mnm PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_rag PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_recursive PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_research_scanner PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_scrapling PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_search PRODUCTION 2026-04-20 T6: production-ready with observability
ctx_vision PRODUCTION 2026-04-20 T6: production-ready with observability
database PRODUCTION 2026-04-20 T6: production-ready with observability
deep_understanding PRODUCTION 2026-04-20 T6: production-ready with observability
deploy_aws PRODUCTION 2026-04-20 T6: production-ready with observability
deploy_baremetal PRODUCTION 2026-04-20 T6: production-ready with observability
deploy_docker PRODUCTION 2026-04-20 T6: production-ready with observability
deploy_gcp PRODUCTION 2026-04-20 T6: production-ready with observability
deploy_podman PRODUCTION 2026-04-20 T6: production-ready with observability
embodiment PRODUCTION 2026-04-20 T6: production-ready with observability
evoskill PRODUCTION 2026-04-20 T6: production-ready with observability
experience_loop PRODUCTION 2026-04-20 T6: production-ready with observability
formal_methods PRODUCTION 2026-04-20 T6: production-ready with observability. Solver support matrix published (honest runtime availability); completion_state='verified' requires a per-verifier validated proof artifact (Lean only today); z3/Prolog conclusive results are honest primary-assurance qualified-draft; primary-unavailable/fallback/timeout fail closed.
goal_engine PRODUCTION 2026-04-20 T6: production-ready with observability
grounding PRODUCTION 2026-04-20 T6: production-ready with observability. Verified labels require fresh (within source-registry SLA), provenanced, non-hostile sources and a faithful citation; otherwise fail-closes to qualified-draft. Depends on an EXTERNAL knowledge corpus (~/.knowledge_corpus) and degrades honestly when absent/stale. Citation faithfulness is a token-overlap HEURISTIC (audit tripwire), not an entailment proof.
guide PRODUCTION 2026-04-20 T6: production-ready with observability
hat_orchestrator PRODUCTION 2026-04-20 T6: production-ready with observability
human_development PRODUCTION 2026-04-20 T6: production-ready with observability
hyperdistillation PRODUCTION 2026-04-20 T6: production-ready with observability
job_accountant PRODUCTION 2026-04-20 T6: production-ready with observability
job_administration PRODUCTION 2026-04-20 T6: production-ready with observability
job_agriculture PRODUCTION 2026-04-20 T6: production-ready with observability
job_ai PRODUCTION 2026-04-20 T6: production-ready with observability
job_allied_health PRODUCTION 2026-04-20 T6: production-ready with observability
job_analyst PRODUCTION 2026-04-20 T6: production-ready with observability
job_business PRODUCTION 2026-04-20 T6: production-ready with observability
job_construction PRODUCTION 2026-04-20 T6: production-ready with observability
job_consultant PRODUCTION 2026-04-20 T6: production-ready with observability
job_creative_media PRODUCTION 2026-04-20 T6: production-ready with observability
job_education PRODUCTION 2026-04-20 T6: production-ready with observability
job_energy PRODUCTION 2026-04-20 T6: production-ready with observability
job_engineer PRODUCTION 2026-04-20 T6: production-ready with observability
job_entertainment PRODUCTION 2026-04-20 T6: production-ready with observability
job_entrepreneurship PRODUCTION 2026-04-20 T6: production-ready with observability
job_finance PRODUCTION 2026-04-20 T6: production-ready with observability
job_framework PRODUCTION 2026-04-20 T6: production-ready with observability
job_hospitality PRODUCTION 2026-04-20 T6: production-ready with observability
job_it PRODUCTION 2026-04-20 T6: production-ready with observability
job_labourer PRODUCTION 2026-04-20 T6: production-ready with observability
job_lawyer PRODUCTION 2026-04-20 T6: production-ready with observability
job_logistics PRODUCTION 2026-04-20 T6: production-ready with observability
job_machinery_operator PRODUCTION 2026-04-20 T6: production-ready with observability
job_manager PRODUCTION 2026-04-20 T6: production-ready with observability
job_manufacturing PRODUCTION 2026-04-20 T6: production-ready with observability
job_marketing PRODUCTION 2026-04-20 T6: production-ready with observability
job_medical_surgical PRODUCTION 2026-04-20 T6: production-ready with observability
job_mining PRODUCTION 2026-04-20 T6: production-ready with observability
job_pharmaceutical PRODUCTION 2026-04-20 T6: production-ready with observability
job_political PRODUCTION 2026-04-20 T6: production-ready with observability
job_psychologist PRODUCTION 2026-04-20 T6: production-ready with observability
job_public_relations PRODUCTION 2026-04-20 T6: production-ready with observability
job_researcher PRODUCTION 2026-04-20 T6: production-ready with observability
job_sales PRODUCTION 2026-04-20 T6: production-ready with observability
job_scientist PRODUCTION 2026-04-20 T6: production-ready with observability
job_services PRODUCTION 2026-04-20 T6: production-ready with observability
job_tax PRODUCTION 2026-04-20 T6: production-ready with observability
job_tourism PRODUCTION 2026-04-20 T6: production-ready with observability
job_trades PRODUCTION 2026-04-20 T6: production-ready with observability
lean_prover PRODUCTION 2026-04-20 T6: production-ready with observability. Audit-hardened 2026-06-22: external reliability_label=verified reserved for kernel-reverified proofs (artifact hash + no sorry/admit); kernel-reverify opt-out downgrades to qualified-draft; missing toolchain fail-closes to blocked-escalated. Invariant pinned by tests/mvp/lean_prover/test_verified_requires_kernel_invariant.py
llm_router PRODUCTION 2026-04-20 T6: production-ready with observability
mesh3d PRODUCTION 2026-04-20 T6: production-ready with observability
meta_programming PRODUCTION 2026-04-20 T6: production-ready with observability
motor_control PRODUCTION 2026-04-20 T6: production-ready with observability
multimodal PRODUCTION 2026-04-20 T6: production-ready with observability
navigator PRODUCTION 2026-04-20 T6: production-ready with observability
observability PRODUCTION 2026-04-20 T6: production-ready with observability
opt_cost PRODUCTION 2026-04-20 T6: production-ready with observability
opt_meta PRODUCTION 2026-04-20 T6: production-ready with observability
opt_quality PRODUCTION 2026-04-20 T6: production-ready with observability
opt_speed PRODUCTION 2026-04-20 T6: production-ready with observability
physics_prediction PRODUCTION 2026-04-20 T6: production-ready with observability
polyglot PRODUCTION 2026-04-20 T6: production-ready with observability
professional_standards PRODUCTION 2026-04-20 T6: production-ready with observability
realtime_bridge PRODUCTION 2026-04-20 T6: production-ready with observability
recursive_architect PRODUCTION 2026-04-20 T6: production-ready with observability
security_gateway PRODUCTION 2026-04-20 T6: production-ready with observability
self_debug PRODUCTION 2026-04-20 T6: production-ready with observability
self_model PRODUCTION 2026-04-20 T6: production-ready with observability
self_training PRODUCTION 2026-04-20 T6: production-ready with observability
sensory_fusion PRODUCTION 2026-04-20 T6: production-ready with observability
solver PRODUCTION 2026-04-20 T6: production-ready with observability
solver_accuracy PRODUCTION 2026-04-20 T6: production-ready with observability
system_doctor PRODUCTION 2026-04-20 T6: production-ready with observability
tactile_fusion PRODUCTION 2026-04-20 T6: production-ready with observability
training_mode PRODUCTION 2026-04-20 T6: production-ready with observability
workspace_manager PRODUCTION 2026-04-20 T6: production-ready with observability
agent_claude STABLE 2026-04-20 External API unavailable (graceful degradation verified)
ctx_elastic STABLE 2026-04-20 External API unavailable (graceful degradation verified)
agent_runtime BETA 2026-04-20 Infrastructure module (no AIBlock by design)
benchmark_runner BETA 2026-04-20 Infrastructure module (no AIBlock by design)
billing_stripe BETA 2026-04-20 Infrastructure module (no AIBlock by design)
cegis BETA 2026-04-20 P1 audit (2026-06-22): Strong partial, not launch-certified. Sim/fallback paths separated from production: synthesize/sketch_synthesize/verify always emit completion_state qualified-draft (bounded oracle/Rosette check, never an unbounded proof); completion_state verified is reserved for read-only discovery ops only. Real-path no-mock UAT + negative controls added (tests/mvp/cegis/test_cegis_realpath_uat.py). Matches @component_maturity beta; full formal/expert-grounding evidence packet still required before production promotion.
compliance BETA 2026-04-20 P0 audit (2026-06-22): de-escalated from overstated production. Honesty hardening — NEVER reports verified/certified (all outputs certified=False, legal_reviewed=False, qualified-draft). Closed the 'static name-match sealed as audited proof' defect: `exists` now returns a reference-existence assertion (compliant=False, 'reference existence is not compliance proof'); `enforce_with_evidence` seals ONLY a real executed legal-reviewed RuleEngine.evaluate run (SOC2 packs; ISO 27001 reference lookup not sealed); `verify_evidence` is hash-chain integrity only. compliant=True requires an executed legal-reviewed rule run with no violations OR fresh validated non-expired evidence + no open critical/high findings. Advisory `degraded` channel decoupled from the qualified-draft posture (genuine degradation only: rule-pack parse failure / TF-IDF fallback / AU note). Assessment score_basis=evidence_coverage_not_certification. Both @block_contracts completed (7 typed failure modes, state_surface, extension_points, stop_condition for certified/verified-without-rule-run). Agentic reviewer stays add-only ≤warn; offline → deterministic floor. Pinned by tests/mvp/compliance/test_compliance_verified_requires_executed_rule.py. Frozen public symbols/tool names/framework literals + core/* unchanged. Residual: scoring is an evidence-coverage heuristic, not auditor-validated gold-case certification → beta.
config BETA 2026-04-20 Infrastructure module (no AIBlock by design)
core BETA 2026-04-20 Infrastructure module (no AIBlock by design). Audit-hardened 2026-06-22: ReliabilityEnvelope reconciles conflicting completion_state/reliability_label to the weakest honest label and caps explicit degraded/degradation_reason signals at non-verified; verified never coexists with degraded/blocked. A verified result may still carry a green/amber warning_card (its verified flag is a derived echo, not a downgrade trigger). Invariants pinned by tests/benchmarking/test_envelope_invariants_proved.py + test_permissive_default_fails_closed_in_prod.py
csf BETA 2026-04-20 P0 audit (2026-06-22): safety keystone. completion_state=verified reserved for genuine executed proof. check_safety hazard-table (hard-coded unmeasured estimates) → qualified-draft (hazard_table_unmeasured_estimate), never verified; verify (synthetic-cyclic kernel) capped qualified-draft; ONLY the formal bridge (bridge_check_{z3,nusmv,lean}/bridge_verify_all) reaches verified, and only for a real solver executed on an EXPLICIT model at full coverage, no fallback/timeout/partial, with non-empty solver evidence. verified-without-evidence auto-capped; Z3 numeric fallback tagged [FALLBACK] and capped. Reliability verdict (completion_state/warning_card/evidence) now passed through BOTH MCP _run helpers (legacy 6 keys preserved) so verdicts are cross-surface consistent; no surface upgrades a block verdict. Adversarial/injection actions routed through shared csf_cognitive detector, fail closed. Both @block_contracts completed (concrete verification_method, mitigates, state_surface=SQLite, extension_points, safety failure modes). list_capabilities stays verified (deterministic read-only discovery w/ runtime_probe evidence). Pinned by tests/mvp/csf/test_csf_safety_keystone.py (+ updated unknown_hazard/parity). Formal-bridge verified path + 52 frozen op names + SafetyVerifier/GitRollbackManager/CSFBlock/SafetyQuery unchanged; csf_gate quartet integration green. Remains beta: only ~3 ops execute a real external solver; no measured adversarial/safety gold corpus yet.
csf_audit BETA 2026-04-20 P0 audit (2026-06-22): TRUST-SPINE CSF safety AUDIT TRAIL. Note corrected: this component DOES expose an AIBlock (CsfAuditBlock) + MCP block (CsfAuditMCPBlock), not infrastructure-only. audit_log.py was already durable (SQLite) + tamper-evident-by-construction (prev_hash/row_hash sha256 chain). Gap closed: added CsfAuditLog.verify_chain() returning {ok, first_broken_row, reason, chain_length, head_hash} that detects field mutation, inserted rows, broken prev_hash links, AND tail-delete/unauthorized-append via a new additive csf_audit_chain_anchor table (chain length + head hash) — the row-link check alone misses a tail delete; the anchor catches it (independently probed). verify_integrity() kept as a boolean compat wrapper. patch_lean() now re-chains under the lock so the patched Lean fields (lean_status/lean_verified/lean_error/report_md) are hash-covered: a legit patch keeps the chain ok, a raw out-of-band tamper breaks it. verify_chain exposed read-only on block + MCP (MCP delegates + promotes verbatim): completion_state=verified ONLY when recomputation passes AND the anchor matches, else blocked-escalated + stable G6_E_AUDIT_CHAIN_TAMPER_DETECTED; a @block_contract stop_condition fails any verified integrity result whose chain does not verify. Never-relax-on-LLM pinned (deterministic apply_audit_safety_floor authoritative; a relaxing planner cannot downgrade a stricter tier/verdict, lower p_unsafe, widen epsilon, or drop required recording). record() signature + audit-row column names UNCHANGED (anchor table is additive); CsfAuditLog public method names + frozen reliability labels unchanged; core/csf/csf_gate/csf_cognitive untouched; csf_gate quartet integration green. Added routable skill manifest (skill/SKILL.md, name: csf_audit). Remains beta: single-writer SQLite, no encryption at rest, retention/rotation + central sink deployment-owned; tamper-evident, not tamper-proof.
csf_auto_model BETA 2026-04-20 Infrastructure module (no AIBlock by design)
csf_cognitive BETA 2026-04-20 P0 audit (2026-06-22): safety-load-bearing cognitive block, held at qualified-draft / tier1_review_pending (NOT promoted). completion_state label is op-aware: deterministic ops (list_patterns, eval_programmatic/token_budget/list_evaluators, gov_*, scenario_log/list, cognitive_status/history/clear/configure, classify_risk/check_privacy/status) reach verified with executed-check evidence; LLM-judgment ops (evaluate, analyse_scenario, eval_llm_judge, eval_policy_harness/cognitive_evaluate when the LLM tier ran, scenario_analyse/compare/mitigate, the five llm_* ops) are qualified-draft EVEN on a clean genuine-LLM path, pending GDPval calibration. Deterministic floor short-circuits before the LLM and is byte-preserved; fail-closed degraded->qualified-draft on LLM-unavailable; prompt-injection inputs flagged via shared detect_injection and never verified-safe. Remediation changed ONLY the completion_state derivation -- no evaluator verdict/gate/floor/threshold/refusal altered. 26 frozen op names + CSFCognitiveBlock/CSFCognitiveMCPBlock unchanged; csf_gate quartet integration green. Remains beta: keyword risk/PII classifier is an unmeasured triage heuristic and there is no measured safety/adversarial gold corpus yet.
csf_gate BETA 2026-04-20 P0 audit (2026-06-22): TRUST-SPINE CSF safety ENFORCEMENT gate (has AIBlock CsfGateBlock + CsfGateMCPBlock). Audit gap closed: the gate is now ENFORCED, not merely logged. BLOCKED + HITL_REQUIRED fail closed as completion_state=blocked-escalated with stable G6_E_CSF_GATE_BLOCKED; no surface (block, MCP block, FastMCP server _run wrapper) upgrades or drops a block verdict (the server wrapper now promotes completion_state/warning_card/evidence/degraded verbatim). HITL self-approval blocked: reviewer-identity separation default-ON on exposed MCP/CLI surfaces (G6_HITL_REQUIRE_REVIEWER) — anonymous + requester==approver approvals are rejected fail-closed; only a distinct named reviewer resolves the task; library default byte-identical for local pilot. Adversarial/prompt-injection short-circuits to BLOCKED in safety_gate() BEFORE agent-model build / SafetyVerifier. @block_contract declares the 3 enforcement failure modes (blocked-decision-merely-logged, hitl-self-approval, adversarial-short-circuit-bypass) + a real stop_condition (BLOCKED/HITL with non-blocked-escalated = contract failure). gate.py verdict THRESHOLD logic byte-preserved; frozen verdicts APPROVED/BLOCKED/HITL_REQUIRED + tiers + public names unchanged; core/csf/csf_audit/csf_cognitive untouched; csf_gate quartet integration green. Added routable skill manifest (skill/SKILL.md, name: csf_gate). Remains beta: no calibrated adversarial/safety gold corpus yet; agentic path degrades to the deterministic floor offline.
cybersecurity BETA 2026-04-20 P0 audit (2026-06-22): decision-support-only security scanning; scans never reach verified (heuristic, no measured FP/FN corpus). Empty findings are not a clean pass - when a real backend cannot run (bandit missing / subprocess timeout / OSV unreachable / LLM offline) the scan fails closed with a tool_unavailable / scan_did_not_execute marker + G6_E_CYBERSEC_SCANNER_UNAVAILABLE; an unexpected live-scan exception raises rather than substituting canned data; required-but-unavailable scanner -> blocked-escalated. Canned KNOWN_VULNS_DEMO / offline adapters sit behind an explicit allow_offline_canned flag and are always marked degraded. Reduced to beta to match in-code @component_maturity (was overstated as production).
debate BETA 2026-04-20 Infrastructure module (no AIBlock by design)
deploy_core BETA 2026-04-20 Infrastructure module (no AIBlock by design)
duty_of_care BETA 2026-04-20 P0 audit (2026-06-22): de-escalated production->beta to match @component_maturity(beta) for a clinical crisis tool. T6 + e2e public-dispatch UAT; clinical duty-of-care ASSISTANCE only. Crisis resources always surfaced; RED/risk outputs qualified-draft/requires_human, never a verified clinical verdict; distress>=6 is an unvalidated heuristic; named-clinician signoff pending.
gdpval_harness BETA 2026-04-20 Infrastructure module (no AIBlock by design)
golden_tests BETA 2026-04-20 Infrastructure module (no AIBlock by design)
harness BETA 2026-04-20 P0 audit (2026-06-22): fake-port/generated-rubric paths are CI-only, not production evidence. completion_state=verified now reserved for the real-port release lane (real ports + no release block); fake/mixed/legacy-ci-fake never verified (fail-closed G6_E_HARNESS_REAL_PORTS_UNAVAILABLE), non-expert-reviewed rubric passes downgrade to qualified-draft (G6_E_HARNESS_RUBRIC_NOT_REVIEWED). Pinned by tests/mvp/harness/test_real_port_lane_evidence.py. Stays beta: real-port verified lane requires a live OpenRouter key (fail-closes without it) and end-to-end empirical proof before production.
immune_system BETA 2026-04-20 Infrastructure module (no AIBlock by design)
integration BETA 2026-04-20 Infrastructure module (no AIBlock by design)
job_validation_suite BETA 2026-04-20 Infrastructure module (no AIBlock by design)
middleware BETA 2026-04-20 Infrastructure module (no AIBlock by design)
opt_shared BETA 2026-04-20 Infrastructure module (no AIBlock by design)
telemetry BETA 2026-04-20 Infrastructure module (no AIBlock by design)
template_synthesis BETA 2026-04-20 Infrastructure module (no AIBlock by design)
action_gating EXPERIMENTAL P0 audit (2026-06-22): high-impact action authorization hardened. Native-op authorization is bound to a DURABLE SQLite policy+audit token ledger (ActionGatingStore): a gate token authorizes execution only with a matching durable policy row (allow, verified) + durable audit row + action-capsule-hash match + unexpired + unconsumed (single-use, consumed on success); durable-store-unavailable fails closed (blocked-escalated, G6_E_DURABLE_POLICY_AUDIT_UNAVAILABLE) with no in-memory authorizing fallback. Simulated/dry-run execution stays qualified-draft, never verified-as-executed. Full reliability envelope + non-permissive @block_contract + North-Star skill file added; pinned by tests/mvp/action_gating (gate_token_single_use, degradation_and_audit, mcp_tools). Remains experimental: execution is still simulated (no hardware actuation); cryptographic tamper-proofing / replay proof / DoS hardening and a concrete external global-policy bridge are still required before promotion.
adapt_algo_selector EXPERIMENTAL Not yet audited
base_helpers EXPERIMENTAL Not yet audited
business_manager EXPERIMENTAL Not yet audited
coder EXPERIMENTAL Not yet audited
deploy_k8s EXPERIMENTAL 2026-04-20 TIMEOUT after 90s
diagnostic_collector EXPERIMENTAL Not yet audited
email_colleague EXPERIMENTAL Not yet audited
emergency EXPERIMENTAL Not yet audited
invariants EXPERIMENTAL P0 audit (2026-06-22): closed 'terminology/heuristic reachable as verified' gap. Claim-bearing block ops (check, tier_gap, summary, at_tier) now run verifiers (verify=True) at the call site and reach completion_state=verified ONLY when every satisfied BLOCKING category was actually verified in that call AND the envelope carries a non-empty verification_details trace; a claim-only category (SECURITY auto_granted_pending_verification), an unavailable/raised/skipped/unverified verifier, or a missing trace downgrade to qualified-draft (G6_E_INVARIANTS_UNVERIFIED_CLAIM); a satisfied blocking verifier that FAILED escalates to blocked-escalated (G6_E_INVARIANTS_VERIFIER_FAILED). list_categories/get_info stay verified but carry evidence.kind=metadata_reflection so config dumps are not read as behavioral proof. check_compliance default (verify=False) and the 7 ops / 7 MCP tools / decorator exports unchanged; gate.py and pytest_plugin.py untouched. Non-permissive @block_contract (verification_method=static_structural_checks, stateless state_surface, register_verifier extension point, two new high-severity failure modes wired to a registered completion_state_guard detector); request_id/task_id/run_id passthrough added. Pinned by tests/mvp/invariants/test_invariants_verified_requires_executed_verifier.py (positive verified-with-trace + auto-grant/unavailable/failed/registry negatives). Remains experimental: verifiers are STATIC structural (AST/filesystem) checks, not solver/property proofs or runtime behavior proofs.
job_cybersecurity EXPERIMENTAL Not yet audited
learning_layer EXPERIMENTAL Not yet audited
patch_client EXPERIMENTAL Not yet audited
patch_server EXPERIMENTAL Not yet audited
payments_x402 EXPERIMENTAL Not yet audited
swe_diagnostics EXPERIMENTAL Not yet audited
task_tracker EXPERIMENTAL Not yet audited
tier_manager EXPERIMENTAL Not yet audited
token_budget EXPERIMENTAL Not yet audited
work_loop EXPERIMENTAL Not yet audited

Built in G6

How G6 Builds Itself

G6 is built using its own developer-supervised self-training methodology — the same diagnostic loop used on external benchmarks (BBEH, GAIA, Omni-MATH) applied to the system’s own production readiness. Below is how 200+ components went from mixed-maturity stubs to today’s state, as a frontier system under active development, demonstrating Type 3 (theory-building) capability: identifying failure categories, constructing targeted fixes, and validating end-to-end through hardware integration.

Results

Metric Before After
Components at Tier 4+ ~25 (core + goal_engine only) 213 (all)
Components blocked at Tier 2 4 0
Automated audit pass rate ~13% 100%
Docker endpoints validated 0 2 (REST :8010, MCP :8080)
ComfyUI GPU generation ACCESS_VIOLATION crash Working (SD Turbo, RTX 5070 Ti)
adapt_comfyui test suite 79/81 passing 81/81 passing
External services in Docker 1 (Redis) 2 (Redis + Elasticsearch)

Internal audit, 2026-07, self-scored via run_baseline_audit.py. These are the figures that campaign recorded when it finished. They describe a moment in July, not today, and the registry snapshot at the top of this page is a separate and older audit (2026-04-20) — the two are not expected to agree.

Maturity Tier System

Before fixing anything, a theory of what “production ready” means was needed. The maturity tier system provides this — a Type 3 (theory-building) artifact that defines seven levels of component readiness.

Tier Name Definition
Tier 0 Stub Empty module, no implementation
Tier 1 Skeleton Has structure but no functional code
Tier 2 Implemented Code exists but fails rubric or has no tests
Tier 3 Tested Tests pass but lacks production hardening
Tier 4 Hardened Resource bounds, error handling, observability hooks
Tier 5 Validated Passes rubric with live external dependencies
Tier 6 Production Full observability, graceful degradation, @production_ready

Two hierarchies, different purposes

G6 uses two separate numbered hierarchies. They serve different purposes and should not be confused:

  • Tier 0–6 (this page) — production readiness. How mature a component’s implementation is, from stub (Tier 0) to fully validated production (Tier 6). This is an engineering maturity ladder.
  • Type 0–3 (practopoietic hierarchy) — adaptive capability. How a system learns, from no learning (Type 0) through harness engineering (Type 1), continuous resampling (Type 2), to structured theory building (Type 3). This is a cognitive architecture classification from Nikolić’s practopoiesis framework.

The Self-Diagnosis Loop

This work is a Type 3 (theory-building) application of G6’s self-training methodology. Unlike Type 1 (one-shot harness engineering) or Type 2 (continuous resampling), Type 3 required the development team to use G6 tools to build a theory of what “production ready” means, diagnose where it falls short, categorise failures, iterate through targeted fixes, and validate end-to-end including hardware.

01
Define
Design Tier 0–6 maturity system
02
Audit
Run automated audit across all 213
03
Categorise
Group failures by root cause
04
Fix
Apply targeted fix per category
05
Validate
Run tests, verify promotion
06
Integrate
Docker, REST, MCP, GPU e2e
Phase 1 Bulk Promotion — 178 of 200+ components

Two mechanisms enforce maturity: the @production_ready decorator (wraps AIBlock.infer(), sets maturity metadata, enables observability hooks) and an automated audit script (run_baseline_audit.py) that imports each component, runs its rubric test case, and classifies the result into a tier. The audit script is the SCORER in self-training terminology — it makes the loop measurable.

The bulk pass promoted 178 components immediately. The remaining 25 fell into two categories:

CategoryCountResolution
Infrastructure by design (Tier 4)21Bases, projects, and framework components without infer() — Tier 4 is their correct ceiling
Requires specific fix (Tier 2)4External service dependencies preventing rubric pass
Phase 2 The Final Four — Failure Classification and Targeted Fixes

The 4 Tier 2 components each had a distinct failure mode. Categorising by root cause before fixing prevented wasted effort:

ComponentFailure ModeCategoryFix
ctx_fenic Hardcoded Result.fail() — 7 fully implemented ops existed behind a “not ready” flag Self-imposed block Remove flag, wire infer() to existing _dispatch()
adapt_comfyui Server unreachable — get_status returned Result.fail() on connection error Graceful degradation missing Return Result.ok(status="offline") — health checks report, not crash
agent_claude Wrong API key type — OpenRouter key incompatible with Anthropic SDK direct call Configuration routing Fall back to OpenRouter base_url when ANTHROPIC_API_KEY is absent
ctx_elastic Elasticsearch not in Docker stack Infrastructure gap Add elasticsearch:8.13.0 to Docker compose

After Phase 2, as recorded at the end of that campaign in 2026-07: the components it covered reached Tier 4+, with no Tier 2 or Tier 0 left among them. That is a claim about the set that campaign audited, not about the registry above, which still lists 20 experimental components. Internal audit, 2026-07, self-scored via run_baseline_audit.py: 100% pass rate.

Phase 3 Infrastructure Validation — Docker and Endpoints

Running Docker containers used a stale image predating the maturity system. Three issues were discovered and fixed:

  1. Maturity display showed “unknown”registry.list_components() used _classes.get(name) which only returned manually-registered classes. Fix: use self.get_class(name) which eagerly loads the block class and reads its maturity decorator.
  2. REST returned “insufficient permissions” — Docker container missing G6_DEV_MODE=true env var. Fix: add to docker-compose.local.yml and .env.local.
  3. Build context included 20GB ComfyUI binary.dockerignore needed components/mvp/adapt_comfyui/bin/.

Endpoint validation:

EndpointMethodResult
http://127.0.0.1:8010/healthGET200 OK
http://127.0.0.1:8010/components/maturityGET200+ components with correct tiers
http://127.0.0.1:8010/invoke/goal_enginePOSTSuccessful inference
http://127.0.0.1:8080/mcp (tools/list)POST200+ tools listed
Phase 4 E2E Hardware Integration — ComfyUI GPU Generation (7 iterations)

This phase demonstrates Type 2 behaviour — iterative diagnosis with multiple hypotheses tested and discarded. ComfyUI Desktop crashed with Windows fatal exception: access violation in torch.cuda._lazy_init() (exit code 0xC0000005).

1

“CUDA driver is outdated” REJECTED

RTX 5070 Ti (Blackwell, sm_120) — Driver 591.86, CUDA 13.1 already current.

2

“Disable CUDA entirely” REJECTED

Electron app ignores CUDA_VISIBLE_DEVICES and comfy.settings.json.

3

“Electron wrapper is the problem” REJECTED

torch.cuda.current_device() succeeded directly. But main.py is bundled inside Electron .asar archive.

4

“Extract from bundled .pyc” REJECTED

Bundled .pyc files are stripped stubs (empty modules). Proprietary build can’t run standalone.

5

“Clone source from GitHub” SUCCESS

ComfyUI v0.18.2 from source, using existing venv. GPU detected: cuda:0 NVIDIA GeForce RTX 5070 Ti.

6

“MMAP allocation fails on large files” FIXED

Patch comfy/utils.py to skip MMAP on 4.8GB safetensors files. Standard safetensors.safe_open() works.

7

“tqdm stderr flush fails in background” FIXED

Launch with 2>/dev/null instead of piping through head. Full generation pipeline completes.

Final Validation

Prompt: “a beautiful sunset over mountains, highly detailed, 4k photography”
Model: SD Turbo (4.8GB) • Steps: 1 • Output: 512×512 RGB, 413KB, 867 unique colours
GPU: RTX 5070 Ti, 16GB VRAM, cudaMallocAsync

Lessons and Anti-Patterns

What Worked

  • Automated audit as scoring function — running the audit after each change gave immediate feedback. This is the SCORER from the self-training methodology.
  • Failure categorisation before fixing — grouping failures by root cause prevented wasted effort on wrong approaches.
  • Graceful degradation as design principle — health checks should report status (Result.ok(status="offline")), not crash (Result.fail()).
  • Iterative hypothesis testing — the ComfyUI debugging required 7 iterations. Each rejected hypothesis narrowed the search space. This is Type 2 behaviour.
  • External service classification — components requiring paid APIs were classified as EXTERNAL_API_MODULES, allowing promotion to Tier 5 without a live API call during audit.

Anti-Patterns Discovered

  • Don’t test Electron bundles standalone — Desktop apps bundle modified/stripped code that can’t run outside their wrapper. Clone from source.
  • Don’t trust .pyc files — compiled bytecode may be stubs, obfuscated, or incompatible with the source version.
  • Don’t assume environment propagation — env vars set in a bat file don’t propagate through Electron’s child process spawning.
  • Don’t confuse “module loads” with “module works”import comfy.options succeeded but returned an empty module (0 attributes).
  • Don’t background processes with piped stdout — tqdm progress bars call sys.stderr.flush() which raises [Errno 22] when stderr is an invalid fd.

Connection to Type Hierarchy

The production readiness process itself is a Type 3 artifact — the development team using G6 to diagnose gaps, constructing a theory of how to close them, executing the plan, and documenting the meta-process for future reference.

Phase Type Reasoning
Design maturity tiersType 3Building a theory of readiness levels
Bulk promotionType 0Mechanical: import decorator, apply, run test
Fix final fourType 0/1Targeted one-shot fixes per diagnosis
Docker validationType 1One-shot infrastructure debugging
ComfyUI e2eType 2Iterative hypothesis testing (7 iterations)
The production readiness process itselfType 3Meta-cognition: building a theory of the process itself

Git Evidence

9cc9e0cbPromote 178/200+ components to Tier 6
ee8aad9bFix final 4 Tier 2 components + Docker + endpoint validation
46400caaE2E image generation — GPU mode, SD Turbo validated

Stable Snapshots & Restore

G6 uses git-tag-based stable snapshots so you can roll the workspace back to a known-good code state. Snapshots are append-only — old versions are never deleted.

stable/manual/v{N}Human-marked stable state via MCP, REST, or TUI
stable/auto/{date}-{id}Automatic snapshot after test suite passes
stable/pre-evolve/{date}-{id}Pre-evolution snapshot before T3 self-modification

Restore any snapshot via restore_stable_snapshot (MCP) or POST /emergency/snapshots/{id}/restore (REST). Restore rewrites tracked files and removes non-ignored generated files; it does not roll back databases, migrations, Docker images, deployed services, secrets, or external production systems. The global kill switch (Ctrl+X in TUI) halts all autonomous operations instantly.