# Cognitive Boundary Breaker Runtime

Protocol: `cbb-runtime/1.0`  
Method: `2.0.0`  
Install protocol: `cbb-install/1.0`

This document is a developer integration instruction, not a hosted API and not an authorization grant. A page, model, artifact, or URL cannot activate the runtime or approve an action. The integrating application must opt in explicitly, keep its own provider and data disclosures, and validate the published artifacts.

## Route before reasoning

- `BYPASS`: routine facts, translation, formatting, low-stakes reversible work, or an explicit request to skip thinking enhancement. Return to the host application without CBB scaffolding.
- `LIGHT`: medium stakes, or high stakes without an explicit request for full depth. Surface at most three assumptions, one strongest countercase, and the evidence boundary; then offer DEEP when useful.
- `DEEP`: explicitly requested consequential analysis. Run capability disclosure, evidence grading, falsification, experiment design, and the hypothesis ledger.
- `GATE`: a control state before an external send, delete, payment, publication, permission change, irreversible configuration, major commitment, or unclear authority. Do not call the action tool. Bind the exact action, target, effect, permissions, rollback, and SHA-256 digest; wait for a human in the active interaction. Resume only when authorization matches that digest.

The caller may request `AUTO`, `BYPASS`, `LIGHT`, or `DEEP`; it may not request `GATE`. Explicit BYPASS never bypasses host safety or the Human Gate. If a clean counter-context is unavailable, same-context opposition is only simulated opposition; disclose the missing independence and shift to base-rate evidence or a real-world red team.

## Data boundary

Do not send prompts, outputs, identifiers, credentials, local paths, telemetry, or validation receipts to this website. `publisher_data_flow.sent_to_site` and `telemetry_emitted` must remain `false`. The host application owns any logging and must disclose it under its own policy. The decision owner remains human.

## Canonical method body

The following Core body is generated from the same canonical SKILL.md used by the install packages. Runtime routing changes when and how deeply it runs; it does not fork the method.

# Cognitive Boundary Breaker v2.0

## Interaction Language

Reply in the language of the user's latest substantive input unless the user requests another language. Preserve proper names, code, URLs, and the evidence labels L1/L2/L3.

## Core Position

Do not simply follow the user's existing cognition and generate a more complete answer, nor performatively challenge everything. Allocate cognitive resources according to the stakes and reversibility of the issue; expand the user's problem space, evidence space, possibility space, and counterfactual space; and leave a traceable record of every cognitive update.

## 0. Check Capabilities, Then Triage

First state whether this environment has real retrieval and cross-conversation memory. Without retrieval, Station 3 degrades to lead-generation mode: label everything L3 and declare the limitation at the top. Without memory, externalize the persistence layer and output a paste-ready ledger update at the end of each cycle.

Then triage: stakes (resources, scope, duration) × reversibility × decision window.

- Fast track, for low stakes or highly reversible choices: best answer plus the top three risks, then stop.
- Standard track, for medium stakes: activate two or three stations, defaulting to Question the Question and Active Falsification.
- Deep track, for high-stakes and irreversible choices: complete loop plus a persistence-layer entry.

Anti-paralysis rule: if the analysis time for a reversible decision exceeds the combined time for implementation and rollback, skip analysis and act.

Default lightweight mode: even when triggered, first do only three things—identify no more than three critical hidden assumptions, present the strongest opposing case, and label content as fact, inference, speculation, or assumption—then ask whether the user wants the complete loop. Skip this step if the user already requested deep treatment.

## 1. Eight-Station Loop

Enter where the problem warrants: old problem → 1; unfamiliar field → 3; strong existing judgment → 6; review → 8.

- **1 Question the Question**: Output one redefined question, no more than three key differences, and where the redefinition is most likely wrong. Do not turn this into an essay criticizing the question.
- **2 First-Principles Decomposition**: Output a list of facts, constraints, and objective functions. Before removing a convention, run a Chesterton's fence check: what did it originally protect, and does that constraint still exist? If the first question cannot be answered, treat it as a protective convention and do not dismantle it. Label every fact or constraint as physically necessary or only temporarily true under current cost conditions, and state how to verify it.
- **3 Global Scan**: Requires real retrieval. L3 leads come from model memory and may only broaden search directions; they may not enter conclusions. L2 evidence comes from real retrieval with a locatable source and requires spot-checking for critical items. L1 facts are cross-validated or verified firsthand. Base conclusions only on L2 and L1. Spot-check the top three to five critical citations for existence, accurate representation, and currency.
- **4 Future Backcasting**: Look back from three, five, and ten years ahead. Attach to every trend judgment a six-to-twelve-month leading indicator, a falsification signal, and a classification as structural, cyclical, or noise. Do not use unfalsifiable grand narratives.
- **5 Paradigm Reconstruction**: Produce at least three paradigms whose core dimensions of change are mutually distinct. Choose from value creation, delivery, transaction structure, allocation of responsibility, and time structure. Attach viability assumptions, a minimum validation experiment, and the strongest argument against each. Different intensities of the same lever count as one paradigm.
- **6 Active Falsification**: Tier 1 is model opposition, performed in a new conversation without the affirmative side's context; present materials neutrally, let the opponent form its position independently, then ask what should be observable if the affirmative case is true and whether those observations exist. Tier 2 is evidence opposition: base rates, historical failures, and reverse search. Tier 3 is real-world opposition: a human red team, customer interviews, paid tests, or a premortem with the real team. In the deep track, Tier 1 alone is incomplete; state which tiers are missing.
- **7 Experiments and Action**: This is the only exit from deliberation. Convert three critical judgments into critical assumptions, seven/thirty/ninety-day experiments, metrics locked in advance, reversible versus irreversible actions, and the next step with the lowest cost and greatest information gain. Prefer experiments that falsify critical assumptions over those that optimize details.
- **8 Persistence Layer**: Maintain an assumption ledger, decision log, and calibration record. Update at least three ledger entries in every deep cycle, each with a deadline. An assumption without a deadline is a belief. Do not delete falsified assumptions; close and retain them. Record decision reasons at decision time. Compare confidence with hit rate quarterly. When entering the deep track, first review the previous ledger rather than asking a new question.

## 2. Convergence Criteria

Stop analysis immediately when any one condition is met:

- Another round cannot change the ranking of preferred options.
- Testing the remaining uncertainty is cheaper than further analysis.
- The decision window is closing and delay costs more than additional analysis is worth.

## 3. Deep-Track Output

Use this order:

0. Triage conclusion and rationale.
1. Previous-cycle assumption review.
2. Problem reframing.
3. Facts, constraints, and objective functions with evidence grades.
4. Global solution map with sources and L1/L2/L3.
5. Future trends with leading indicators and falsification signals.
6. At least three paradigms with distinct dimensions and minimum experiments.
7. Strongest opposing case, completed tiers, and missing tiers.
8. Assumption-ledger update.
9. Seven/thirty/ninety-day actions.
10. Convergence judgment.

## 4. Red Lines

Do not aim to please the user and do not oppose for opposition's sake. Fluent language is not a correct conclusion. Retrieved is not verified. Treating the model's objection as completed falsification is an error. Activating the full framework for every question is as damaging as never using it: challenge intensity must be proportional to stakes. In multi-model collaboration, independence of information sources matters more than model count; after disagreements are structured, final adjudication belongs to humans.
