title: "Gather — "What do we NOT know?"" source: "tasks/TFW-55__canonization_program/research/iter1/2_gather.md"
Gather — "What do we NOT know?"¶
Mindset: Explorer. You're mapping unknown territory. Widen before you narrow. Every assumption is a question. Test: "Can I name every dimension and its alternatives without checking my sources?" Parent: HL-TFW-55 Goal: Map the evidence, unknowns, and genuine decision dimensions behind TFW's identity, self-canon architecture, founder knowledge, and two-surface exposition without selecting a winner.
Dimensions¶
Identify the independent decision factors in this problem. Each dimension represents a degree of freedom — a variable the final solution must choose a value for.
For each dimension, list ≥3 alternatives. Do NOT mark any alternative as "recommended" — all options remain open until Challenge.
| Dimension | Alt A | Alt B | Alt C | Alt D (if any) |
|---|---|---|---|---|
| D1 — Primary category | Philosophy with a method beneath it | Discipline for traceable AI-delegated work | Methodology combining stable behaviors | Practical prompt-and-files framework with philosophical framing |
| D2 — Official authority | Living specification alone | Repository corpus alone | Corpus + canonical essay + living specification, with explicit roles | Separate minimal canon outside the present surfaces |
| D3 — Minimum trace payload | Raw or lightly curated transcript | Provenance chain: actors, activities, inputs, derivations | Selected intent, constraints, decisions, rejected alternatives, result, and continuation state | Full assurance chain: selected trace + evidence + review + verified knowledge |
| D4 — Project self-awareness claim | Retire the term; use inspectability/self-description | Retain the frozen six-question operational capability | Expand to reflexive learning and self-correction | Use the metaphor publicly but define it only in implementation terms |
| D5 — Human/agent relation | AI as tool; all meaningful decisions remain direct human acts | AI as bounded delegate; human owns purpose, authority, acceptance, and stop decisions | AI as team member on an explicit assurance gradient | Role differs by domain/risk; no single canonical relation |
| D6 — Public exposition architecture | Root README is the complete canonical explanation | Root doorway + .tfw/README.md canonical essay |
Living specification first; READMEs are summaries only | Future 20–30 page guide becomes the primary human authority |
| D7 — Adoption/teaching sequence | Explain the complete method/framework first | Establish concepts and vocabulary, then practice | Useful work → experienced limit → next mechanism → named principle | Branch immediately by prior knowledge, work risk, and role |
| D8 — Editions in the foundation | Omit Editions from the essay | Mention as current proportional implementations only | Use Light → Assisted → Full as a non-universal learning example | Make the progression the essay's central explanatory spine |
If fewer than 3 independent dimensions exist, skip this table and use a comparison matrix (pros/cons) in Findings instead.
Findings¶
G1 — Evidence discipline and source boundary¶
The Gather loop read 15 stage-specific local sources, the configured soft ceiling. The frozen TFW-55 HL, root README, KNOWLEDGE.md, conventions, and glossary were already loaded as workflow context and were not re-counted as stage reads. Because unrelated TFW-53 work was changing the shared checkout, the current essay and knowledge/philosophy.md were read from the coordinator's research baseline f131f71 rather than from uncommitted workspace state.
| Source lane | Sources | Epistemic use in this stage |
|---|---|---|
| Current official exposition | .tfw/README.md at f131f71; frozen TFW-55 HL; root README at the same baseline |
Current claims and contradictions, not historical invariants |
| Verified/curated project memory | knowledge/philosophy.md at f131f71; preloaded KNOWLEDGE.md |
Verified facts, one-source beliefs, and decision index kept distinct |
| Earlier TFW reasoning | TFW-25 RES; TFW-32 Phase D RF; TFW-52 Iterations 1–3 RES; Git range ee8d444..bc6779e |
What TFW previously treated as philosophy, positioning, invariant, rejected architecture, and evidence boundary |
| Founder/teaching corpus | INNO-8 HL; INNO-13 curriculum map, assessment model, Day-1 practice, and three verbatim delivery feedback files | Owner claims, designed pedagogy, operational teaching devices, and live founder-led observations—not universal facts |
| Secondary transformation corpus | RF__internal_03_transformation.md from the university work |
A claimed outcome whose causal strength must be checked against its own design |
| Narrow external controls | W3C PROV-O; NIST AI RMF Core | Counterexamples to novelty claims about provenance, responsibility, delegation, oversight, and documentation; not the systematic comparison reserved for Iteration 2 |
One named HL source could not be used directly: tasks/TFW-36__content_marketing_blog_series/ contains no readable files in the active tree. The HL's description of a TFW-36 source-integrity failure remains context, but this stage does not treat inaccessible task content as evidence.
G2 — The current canonical essay contains a real thesis and several implementation-level absolutes¶
The existing .tfw/README.md already supplies a coherent candidate thesis: dialogue loses context; durable files preserve intent, constraints, decisions, and rejected alternatives; new humans or agents can resume; and a discipline is needed beyond a giant prompt or chat export. This is evidence against the strongest claim that TFW is only an arbitrary collection of filenames.
The same essay also collapses levels that TFW-55 must separate:
| Current claim pattern | What the internal corpus supports | What remains overclaimed or contradicted |
|---|---|---|
| “The fundamental bottleneck” is context/knowledge loss | Repeated TFW architecture decisions and teaching tasks address continuity loss | No comparative evidence establishes it as the singular fundamental bottleneck across AI work |
| “Traces Over Code”; code can be regenerated and the same output reproduced | Durable intent and decisions matter across domains; TFW-52 found a portable purpose → bounded work → trace/outcome → continuation payload | Code-centric wording contradicts domain independence; stochastic regeneration and “same output again and again” are not defensible guarantees |
| One deterministic lifecycle and the same ritual in every domain | Full has a structured lifecycle | Light, Assisted, and Full deliberately use different machinery; TFW-52 removed Team and rejected one universal artifact spine |
| Working traces are the documentation; nobody maintains it separately | Work artifacts do produce useful context | Review, evidence collection, consolidation, pruning, status reconciliation, and essay maintenance are real work; “self-maintaining” hides that labor |
| Output requires no manual editing; fix prompt/context instead | No-placeholder and complete-output rules are valid quality aspirations | Human acceptance, review, corrections, and evidence are first-class TFW mechanisms; the absolute erases them |
| RF or generated trace is the source of truth | RF is an important execution record | Evidence, REVIEW, frozen contracts, Git baseline, and verified knowledge now carry distinct and sometimes higher authority |
| AI agents are team members | TFW uses role-bearing agent sessions in real workflows | TFW-52 Iteration 3 and D59 show that another session does not imply identity, permission, independence, or authority |
The essay itself calls TFW both a discipline and a methodology, while the repository and public README also call it a framework. That is not merely wording drift; it is the unresolved D1 decision.
G3 — A stable semantic payload exists, but it is smaller than current Full and larger than provenance alone¶
TFW-25 treated eight README values as the philosophical face of TFW, but its taxonomy was an intentional consolidation of then-current practice, not proof of timeless invariants. Several values—especially Traces Over Code, fixed artifact language, and a deterministic lifecycle—were later narrowed by domain-agnostic Editions and evidence work.
The strongest internal invariant found so far comes from TFW-52 Iteration 1's falsification of its own preferred historical story. Early artifacts did not support a universal goal → Working Backwards → trace → knowledge chain. They supported a name-neutral payload:
purpose / current context
↓
bounded, checkable work
↓
persisted result and trace
↓
reusable continuation context
That payload appears across Full, the four-file Light baseline, Assisted's manual fallback, and the non-code teaching material. It is also consistent with knowledge/philosophy.md F32–F33: simplification preserves goals, sources, decisions, verification, and durable knowledge; teaching should produce useful work and enable independent continuation.
However, W3C PROV-O is a direct counterexample to any claim that recording entities, activities, agents, derivation, delegation, responsibility, revision, and primary sources is unique to TFW. PROV can even express provenance of provenance. Therefore:
- provenance is a necessary comparison category, not TFW's distinct identity by itself;
- TFW's candidate distinction, if one survives Iteration 2, must be in the selection and use of traces to organize ongoing work, bounded delegation, review, and continuation;
- “trace” must not expand until it means every log, transcript, hidden reasoning process, or full audit graph.
The D3 decision remains open between a selected continuity record and a richer assurance chain. Raw transcripts are counterevidence because the current essay already rejects chat exports as unusable linear blobs.
G4 — Human responsibility is central internally, but it is not a novel category claim¶
The internal corpus repeatedly preserves human authority:
knowledge/philosophy.mdF25 says the framework proposes decision infrastructure and the human chooses;- INNO-8 says management analogies are teaching tools, not ontological claims;
- the Day-1 “Ten Decisions” exercise separates collection/organization from judgment and final acceptance;
- the curriculum requires permissions, stop conditions, source checks, and an owner matrix;
- the assessment model treats human authority, source verification, trace, and facilitator independence as critical gates rather than inferred learning.
The narrow external control weakens novelty, not relevance. NIST AI RMF Core already requires documented roles and responsibilities for human–AI configurations, human oversight, executive responsibility, scope, risk, impacts, and periodic review. Its playbook also records go/no-go decisions and organizational accountability. TFW cannot credibly present “humans remain responsible and roles must be explicit” as a new principle unique to TFW.
The open question is whether TFW turns those governance ideas into a distinctive, proportionate daily work method across non-regulated domains. That requires the systematic comparison reserved for Iteration 2. H2 findings in this iteration are therefore provisional.
G5 — “Self-aware project” is internally operationalizable, but the label has no independent validation¶
The frozen six-question definition maps cleanly to existing project capabilities:
| Capability | Internal trace that can answer it | Boundary |
|---|---|---|
| Purpose and why | HL vision, owner insights, frozen baseline | A trace can preserve a stated purpose; it cannot prove the purpose is wise |
| Known, assumed, unknown | RES hypotheses/gaps, knowledge status, risks | Status must remain explicit; completion is not confirmation |
| Decisions, rejections, evidence | task history, RF/REVIEW/EV, Git | The repository contains rejected work as well as current authority |
| Current work state | Task Board + file-based stage state | Only reliable when roles maintain/reconcile it |
| Authority for next decision | role lock, owner verdict, delegation boundary | Files preserve an authority statement; they cannot create legitimate authority |
| Continuation without original chat | compact artifacts, local files, task traces | “No re-explanation” is an aspiration; missing tacit knowledge remains possible |
TFW's own history is strong internal evidence of self-application: it uses research, evidence, rejection, restoration, and new tasks to revise TFW. The TFW-48/49 path is especially important negative evidence. Across 53 commits, the project introduced a large methodology rebaseline and repository-local commit-identity runtime; the restoration commit bc6779e then reset the tracked tree to v0.9.0, deleting 27,103 lines while changing 149 files and preserving the rejected experiments in Git history. This shows that the corpus can retain a rejected path and restore current authority without rewriting history.
It also shows why “the repository is the canon” is insufficiently precise. The repository simultaneously contains current authority, historical authority, rejected experiments, owner corrections, and restoration evidence. A newcomer needs an official selected exposition and current living specification to know which layer governs now.
No independent reader has yet been asked whether “self-aware project” clarifies these capabilities or sounds anthropomorphic. D4 must retain at least the operational term, a non-anthropomorphic replacement, and retirement as live alternatives through Extract.
G6 — The founder/teaching corpus contains conceptual material, pedagogical devices, and outcome claims that must not be merged¶
The teaching sources do contain concepts missing or underdeveloped in the current essay:
| Candidate concept | Provenance classification | Consequence if retained |
|---|---|---|
| AI becomes a working environment, not a single chat or model | Founder/course thesis in INNO-8; repeated in practice | Candidate philosophical framing; not a measured universal fact |
| Prompting is bounded delegation: outcome, context, limits, criteria, verification | Designed method and operational exercises | Strong method-level content; must preserve where human/AI analogy stops |
| Product quality depends on the environment supplied around the model | Demonstrated in a controlled classroom exercise using the same model and four input states | Strong teaching example; not proof that model choice never matters |
| The human should provide purpose/value and acceptance, while an agent may plan tasks | Founder Day-3 formulation | Needs qualification: compatible with bounded delegation only if the human still owns scope, authority, stop, and acceptance |
| A project folder can preserve who we are, decisions, tasks, and progress across sessions | Practice script and Light/Assisted experience | Supports operational self-description and continuation; does not prove lossless memory |
| Philosophy becomes discussable after participants have useful practice | Day-3 observation: unplanned manifesto discussion after the practical sequence | Useful teaching observation; one founder-led group, not causal evidence |
The live course evidence also falsifies a simple success story:
- Day 1 was positively received, yet the Light folder/files segment was rushed and “people did not understand the folder at all.”
- Day 2 allocated more time to practice and the shared starter “worked normally,” but the speaker explicitly noted that tests were not run and Codex did not always follow the process unless invoked clearly.
- Day 3 finished successfully and participants discussed meaning, but the evidence records timing and speaker/owner judgment—not delayed retention, unfamiliar-task transfer, or facilitator independence.
- The university report calls a Day-1→Day-2 increase in concrete artifacts a “mindshift,” but attendance fell from about 90 to 58 and the tasks changed. Its own evidence cannot isolate teaching effect, participant selection, task structure, or AI assistance. The result is a useful observed difference, not causal proof.
- INNO-13's later assessment model is more trustworthy precisely because it marks post-design gates
NOT RUN / DEFERRED, requires individual distributions, delayed recall, near/far transfer, per-device evidence, and forbids causal claims from a small pilot.
Therefore H3 is neither confirmed nor refuted. The corpus contains genuine candidate foundation knowledge, but every item must be classified as invariant, owner interpretation, pedagogical device, field observation, example, or open claim before it enters the essay.
G7 — Editions are evidence for proportionality and learning, not a universal philosophy ladder¶
TFW-52 provides three bounded results relevant to H4:
- Iteration 1 refuted a universal historical spine and narrowed teaching to a guided/adaptive sequence. It explicitly left delayed retrieval and transfer unproven.
- Iteration 2 separated capability from guarantee: hooks available ≠ adapter tested; next-start catch-up ≠ calendar scheduling; recovery ≠ loss prevention; attribution ≠ authentication.
- Iteration 3 found that a stable Team edition was not justified. Separate tasks supply persistence/routing/visibility, not independent identity, permissions, authority, recovery, or measurable advantage. The actual product line is now Light, Assisted, and Full, with Team removed.
The founder-led course supports the pattern “useful work → visible limit → next mechanism,” but also exposes its boundary: when the practical Light folder was introduced too late and too abstractly, it failed; when more practice and a starter were provided, comprehension improved by instructor observation but was not measured. This supports D7 Alt C as a serious configuration and simultaneously blocks treating it as doctrine.
The current root README's “same lifecycle, same artifacts” language directly conflicts with Editions. The essay may use Editions as an example of proportional discipline, an adoption bridge, or omit them in favor of the future guide; Gather does not select among those D8 alternatives.
G8 — Strongest framework-only challenger configuration¶
Per coordinator direction, this configuration remains intact through Challenge even where evidence weakens it.
F-only challenger: TFW is a useful repository-specific integration of docs-as-code, decision/provenance records, staged prompt workflows, AI governance, evidence/review, and knowledge management. Its philosophical vocabulary explains the integration, but it does not create a new fundamental discipline. Its strongest verified value is practical continuity for agent-mediated work, not ontological novelty or a universally valid theory of cognitive delegation.
| Evidence supporting the challenger | Evidence against eliminating the higher-level identity early |
|---|---|
| W3C PROV covers provenance, responsibility, delegation, derivation, revision, and primary-source links | TFW's stable payload is about selecting and using traces to continue bounded work, not only representing provenance |
| NIST AI RMF covers documented human-AI roles, oversight, accountability, and review | TFW attempts a proportionate daily method across code, research, education, documents, and business work rather than only AI-risk governance |
| Current behavior depends heavily on repo files, templates, role prompts, and artifact names | Light/Assisted preserve a smaller semantic payload when Full mechanics disappear, suggesting a method above one implementation |
| TFW-25 borrowed a conventional values/principles/rules taxonomy; TFW-32 positioned against familiar knowledge tools | The repository uses the method to reject, restore, and revise itself; the self-application is observable, not merely a brand claim |
| The 27,103-line TFW-48/49 rollback shows self-redesign can generate bureaucracy and false authority | Preserved negative trace made that rollback and minimal TFW-50 correction possible without erasing history |
| Teaching feedback is founder-led and does not prove transfer or independence | Operational exercises do instantiate bounded delegation, source checking, human acceptance, and continuation outside software engineering |
The challenger is a full D1/D2 configuration, not a rhetorical objection: practical methodology/framework + layered repository authority + selected trace payload + bounded human oversight + proportional Editions as optional teaching aids. Extract must cross it against the philosophy/discipline configurations without stripping away its strongest form.
G9 — Contradiction and overclaim map for Extract¶
| Conflict | Side A | Side B | Current disposition |
|---|---|---|---|
| Category | Essay: discipline + methodology; public repo: framework | Owner thesis: deeper philosophical foundation | Open D1; Iteration 2 needed for adjacent-practice boundary |
| Canon authority | Repository is primary corpus and self-applying record | Rejected/history/current rules coexist; current essay contains false absolutes | Open D2; corpus alone cannot select official current meaning |
| Trace | Selected intent/decision continuity | Provenance, evidence, review, verified knowledge are separate and richer | Open D3; do not equate trace with transcript or full assurance |
| Self-awareness | Frozen six-question capability | Anthropomorphic reception untested; files cannot create knowledge or authority by themselves | Open D4; operational capability survives, public label unresolved |
| Human role | Founder: give AI the goal, let it determine tasks | TFW: human owns purpose, authority, scope boundaries, acceptance, stop | Potential reconciliation: outcome-led delegation; exact wording unresolved |
| Knowledge generation | Traces generate documentation automatically | Consolidation/review/pruning and drift repair require deliberate work | Narrow automatic claim; distinguish capture by work from maintained knowledge |
| Teaching sequence | Light → Assisted → Full worked in founder-led delivery | Day 1 folder failed when rushed; transfer and independent facilitation untested | Current hypothesis only; adaptive alternatives remain live |
| “Mindshift” | Speaker/university reports successful change | Attrition, changed tasks, no control, delayed transfer not measured | Observation, not causal efficacy claim |
| Brand tagline | “The thinking is the product” captures value of reasoning | “Thinking” can imply hidden chain-of-thought; TFW actually preserves selected inspectable traces | Keep as candidate brand language; semantic fit unresolved |
G10 — OODA orientation and Gather decisions¶
Observe. The internal corpus provides evidence for a portable semantic payload and a self-applying repository, but it also records major reversals, edition differences, weak teaching claims, and older public absolutes. W3C and NIST independently cover much of provenance and human oversight.
Orient. This challenges the initial tendency to frame the decision as “fundamental philosophy or mere framework.” At least eight independent factors can vary. A configuration may retain a philosophical foundation while conceding that traceability and accountability are composed from known practices; another may be a practical methodology without needing a new category.
Decide. Gather is sufficient for focused mode:
- the required external controls were used against two exact novelty claims;
- the Briefing gap was closed by separating stable payload, founder knowledge, teaching evidence, overclaims, and the framework-only challenger;
- eight independent dimensions with at least three alternatives each are defined;
- every H1–H4 conclusion that depends on broad external comparison or independent readers remains explicitly open.
Act. Preserve D1–D8 and the F-only challenger for the Configuration Space in Extract. Do not edit the frozen HL. No amendment-proposal candidate is mature yet: the evidence presently narrows free-section claims or supplies alternatives for later Challenge, rather than demonstrating that a frozen commitment must change.
Checkpoint¶
| Found | Remaining |
|---|---|
| A portable internal payload survives artifact and Edition changes: purpose/context → bounded/checkable work → persisted result/trace → reusable continuation context. | Whether that composition is distinct enough to name a discipline or methodology requires Iteration 2's systematic adjacent-practice comparison. |
| The repository demonstrably preserves current work, rejected paths, restoration, and self-revision. | Whether an independent reader can determine official meaning from the proposed two surfaces is untested. |
| The current essay contains a valid continuity thesis plus deterministic, code-centric, self-maintaining, same-artifact, and unbounded-agent overclaims. | Exact replacement language belongs after Extract/Challenge, not Gather. |
| Founder/teaching sources add delegated-cognition, bounded-delegation, useful-artifact, and problem-led learning concepts. | Founder claims, teaching devices, field observations, and invariants still need configuration-level classification. |
| W3C PROV and NIST AI RMF show provenance and human oversight are not unique TFW inventions. | TFW's possible distinction may be the operational combination; broad novelty judgment is deferred. |
| Eight independent dimensions and the strongest framework-only challenger are live. | Extract must cross-reference them without prematurely selecting the owner-preferred self-canon/fundamental framing. |
Sufficiency: - [x] External source used? — W3C PROV-O and NIST AI RMF, each as a narrow control against a specific internal novelty claim. - [x] Briefing gap closed? — stable payload, claim provenance, founder-knowledge classes, contradictions, Editions evidence, and the full challenger are mapped. - [x] Dimensions identified? — D1–D8, each with at least three alternatives.
Stage complete: YES → User decision: close stage — coordinator accepted the eight dimensions, narrow external controls, provenance separation, open H1–H4 boundaries, and intact framework-only challenger.
Coordinator steering for Extract¶
- Separate
repository = primary corpusfromrepository alone = official authority; each configuration must say what preserves history/evidence and what selects current official meaning. - Apply human-reader versus agent-orientation as a mandatory compatibility test for D6/H4. Promote it to another dimension if it varies independently.
- Use the TFW-48/49 episode only as evidence of authority ambiguity, recoverability, and over-engineering risk—not as evidence by itself for or against a fundamental philosophy.
- Keep H2 provisional until Iteration 2: W3C/NIST refute component novelty, not necessarily distinctiveness of the operational composition.
- Keep teaching observation, proposed pedagogical mechanism, and demonstrated learning outcome as three separate states.
- Carry the strongest framework-only challenger unchanged into the Configuration Space; only Challenge may narrow or eliminate it.
No amendment proposal is mature at the Gather gate.