Skip to content

[Roadmap][Future] Context utility, layered state, and governed learning #265

Description

@kevintseng

目標

把下一版本候選的 context work 組織成一條有證據、可分別作 GO/NO-GO/DEFERRED 決定的 roadmap:證明精選上下文是否改善任務結果、釐清 state/evidence/memory/observation 邊界、量測 relations 的額外價值,並讓重複 friction 只能透過受治理的 improvement proposal 進入產品工作。

這是 future roadmap,不改變目前 release scope;版本號、milestone、日期、assignee 與任何 implementation 都需另行決定。

已存在並直接重用的工作

這些 issue 已涵蓋 context volume、progressive disclosure、task outcome、token/tool-call cost、privacy 與 freshness;本 roadmap 不建立重複版本。

新增的缺口

  1. [Future][Architecture] Define the state, evidence, memory, and observation contract #262:定義 state、evidence、memory 與 latest observation 的單一責任、生命週期與 failure semantics。
  2. [Future][Research] Evaluate whether graph relations improve agent answers beyond text recall #264:延用 [Future][Token Efficiency] Confirm measurable usage and freeze benchmark contract #251[Future][Token Efficiency] Build isolated paired A/B runner and freeze baseline #252,判斷 graph relations 是否比 text recall 額外提升正確率與降低探索成本。
  3. [Future][Governance] Turn recurring agent papercuts into evidence-linked improvement proposals #263:把重複 agent papercut 整理成 evidence-linked improvement proposal,但保留 human accept/reject/defer,禁止自動修改產品或 Skill。

建議順序

#251 measurable contract
  └─ #252 paired runner + baseline
       ├─ #250 compact/progressive-disclosure path (#253/#254/#258)
       └─ #264 relation-aware evaluation

#250 results + #257 provenance/freshness
  └─ inform #262 layered context disposition

#255 observability disposition
  └─ #263 governed papercut proposals (GO or DEFERRED)

此順序只表示 evidence dependency,不代表 implementation 承諾。#262 可先做現況映射,但正式 disposition 應讀取 #250#257 的實際結果。

全域原則

  • 以 task correctness、abstention、false grounding、tool calls、authoritative usage、wall time 與 human correction 衡量價值;不以記憶、edge、citation 或 Skill 數量代替成果。
  • 缺少權威 token usage 時標為 UNMEASURABLE,不得以 bytes/characters 冒充。
  • 精簡上下文不可移除 degraded、truncated、conflict、provenance、freshness 或 scope truth signals。
  • Citation/receipt 只證明其直接事實,不自動證明內容正確、已接受或已執行。
  • 自動化可以提出候選,不得自行接受 proposal、改 task state、改 Skill、改 source 或升格為產品政策。
  • 重用現有 retrieval、ranking、staging、review 與 message lifecycle;禁止為每個研究方向新增平行 store、dashboard 或 approval system。

Roadmap 完成條件

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    enhancementNew feature or request

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions