{"id":"94272519-c9ec-433a-b0ac-914e9a3b9c96","entityType":"agent","slug":"clawhub-zeekr0808-hue-multi-model-consensus","name":"Multi Model Consensus","canonicalUrl":"https://www.xpersona.co/agent/clawhub-zeekr0808-hue-multi-model-consensus","canonicalPath":"/agent/clawhub-zeekr0808-hue-multi-model-consensus","generatedAt":"2026-10-10T21:55:11.741Z","source":"CLAWHUB","claimStatus":"UNCLAIMED","verificationTier":"NONE","summary":{"evidence":{"source":"editorial-content","verified":true,"confidence":"high","updatedAt":"2026-10-10T18:44:47.834Z","emptyReason":null},"description":"多模型决策委员会 — 消除单模型偏见，通过多轮分歧讨论产出客观决策参考。支持3-6个模型同时评审，提供量化投票矩阵和6段式共识报告。触发条件：包含「多模型决策」或「多模型委员会」时自动激活。 Skill: Multi Model Consensus Owner: zeekr0808-hue Summary: 多模型决策委员会 — 消除单模型偏见，通过多轮分歧讨论产出客观决策参考。支持3-6个模型同时评审，提供量化投票矩阵和6段式共识报告。触发条件：包含「多模型决策」或「多模型委员会」时自动激活。 Tags: latest:1.9.1 Version history: v1.8.0 | 2026-05-16T14:47:53.996Z | auto **Summary:** This release enforces stricter process control and standardization for the multi-model decision committee, focusing on compliance, template usage, and evaluation transparency.","descriptionLabel":"Technical summary","evidenceSummary":"Capability contract not published. No trust telemetry is available yet. 1.3K downloads reported by the source. Last updated 10/10/2026.","installCommand":"clawhub skill install s17ag6jke66yamdtr7sz868c5x85ff28:multi-model-consensus","sourceUrl":"https://clawhub.ai/zeekr0808-hue/multi-model-consensus","homepage":"https://clawhub.ai/zeekr0808-hue/skills/multi-model-consensus","primaryLinks":[{"label":"View on ClawHub","url":"https://clawhub.ai/zeekr0808-hue/multi-model-consensus","kind":"source"},{"label":"Homepage","url":"https://clawhub.ai/zeekr0808-hue/skills/multi-model-consensus","kind":"homepage"}],"safetyScore":84,"overallRank":62,"popularityScore":62,"trustScore":null,"claimedByName":null,"isOwner":false,"seoDescription":"多模型决策委员会 — 消除单模型偏见，通过多轮分歧讨论产出客观决策参考。支持3-6个模型同时评审，提供量化投票矩阵和6段式共识报告。触发条件：包含「多模型决策」或「多模型委员会」时自动激活。 Skill: Multi Model Consensus Owner: zeekr0808-hue Summary: 多模型决策"},"coverage":{"evidence":{"source":"public-profile","verified":false,"confidence":"medium","updatedAt":"2026-10-10T18:44:47.834Z","emptyReason":null},"protocols":[{"protocol":"OPENCLEW","label":"OpenClaw","status":"self-declared","notes":"Declared in the public agent profile."}],"capabilities":[],"verifiedCount":0,"selfDeclaredCount":1,"capabilityMatrix":{"rows":[{"key":"OPENCLEW","type":"protocol","support":"unknown","confidenceSource":"profile","notes":"Listed on profile"}],"flattenedTokens":"protocol:OPENCLEW|unknown|profile"}},"adoption":{"evidence":{"source":"CLAWHUB","verified":false,"confidence":"medium","updatedAt":"2026-10-10T18:44:47.834Z","emptyReason":null},"stars":null,"forks":null,"downloads":1292,"packageName":null,"latestVersion":"1.8.0","tractionLabel":"1.3K downloads"},"release":{"evidence":{"source":"CLAWHUB","verified":false,"confidence":"medium","updatedAt":"2026-10-10T18:44:47.834Z","emptyReason":null},"lastUpdatedAt":"2026-10-10T18:44:47.834Z","lastCrawledAt":"2026-10-10T18:44:47.834Z","lastIndexedAt":null,"nextCrawlAt":"2026-10-11T18:44:47.834Z","lastVerifiedAt":null,"highlights":[{"version":"1.8.0","createdAt":"2026-05-16T14:47:53.996Z","changelog":"**Summary:** This release enforces stricter process control and standardization for the multi-model decision committee, focusing on compliance, template usage, and evaluation transparency. - Strongly enforces that, upon evaluator/model changes, all available models must be presented for user selection—auto-selection is prohibited. - Mandates the use of unified, pre-defined templates (`references/OUTPUT_TEMPLATE.md`) for every process step; organizers must not modify or freestyle output formatting. - Introduces and standardizes the required process for model combination recommendations and strict result collection procedures, including mandatory use of `sessions_yield` for subagent result collection. - Narrows the consensus report structure from 6 segments to 5, with clearer labeling and requirements. - Tightens rules on error handling, round progression, and prohibits advancing to next rounds or final report output before all expected results are received or timeout occurs. - Adjusts threshold and evaluation status definitions, listing scoring rules for \"通过\", \"待决策\", and \"有分歧\" with clear icons and consequences.","fileCount":10,"zipByteSize":55400},{"version":"1.7.1","createdAt":"2026-05-15T03:44:09.578Z","changelog":"环境适配矩阵新增 Feishu direct chat 兜底方案说明","fileCount":10,"zipByteSize":54190},{"version":"1.7.0","createdAt":"2026-05-07T04:37:34.639Z","changelog":"V1.7.0: 禁止行为升级为4条铁律；异常处理新增框架绕过类型","fileCount":10,"zipByteSize":54149},{"version":"1.9.1","createdAt":"2026-05-01T00:56:04.986Z","changelog":"删除 version 字段；frontmatter 增加路径约束引用注释，提升审计可见性","fileCount":11,"zipByteSize":55066},{"version":"1.6.9","createdAt":"2026-05-01T00:53:11.543Z","changelog":"删除 version 字段；frontmatter 增加路径约束引用注释，提升审计可见性","fileCount":10,"zipByteSize":53703},{"version":"1.6.8","createdAt":"2026-05-01T00:43:50.846Z","changelog":"删除 version 字段；frontmatter 增加路径约束引用注释，提升审计可见性","fileCount":10,"zipByteSize":53704},{"version":"1.9.0","createdAt":"2026-05-01T00:39:40.393Z","changelog":"删除 version 字段；frontmatter 增加路径约束引用注释，提升审计可见性","fileCount":10,"zipByteSize":53702},{"version":"1.6.7","createdAt":"2026-04-30T09:32:09.750Z","changelog":"- 增强「身份纯净原则」：新增评委子Agent工具隔离要求，子Agent仅可通过 `task` 参数评审，不得调用任何工具（如 spawn、Read、Write、Exec、Search 等），所有外部操作统一由组织者管理。 - 优化第0轮环境兼容性检查流程，明确将其作为系统自动自检以确保运行环境兼容，保障子Agent能正常spawn与回收结果。 - 相关文档（如 SKILL.md、README.md 等）同步修订，反映治理细则和流程说明的新要求。 - 其余决策流程、判定规则与核心报告结构维持原有设计。","fileCount":10,"zipByteSize":53686}]},"execution":{"evidence":{"source":"CLAWHUB","verified":false,"confidence":"low","updatedAt":null,"emptyReason":"No published capability contract is available yet."},"installCommand":"clawhub skill install s17ag6jke66yamdtr7sz868c5x85ff28:multi-model-consensus","setupComplexity":"low","setupSteps":["Setup complexity is classified as HIGH. You must provision dedicated cloud infrastructure or an isolated VM. Do not run this directly on your local workstation.","Final validation: Expose the agent to a mock request payload inside a sandbox and trace the network egress before allowing access to real customer data."],"contract":{"contractStatus":"missing","authModes":[],"requires":[],"forbidden":[],"supportsMcp":false,"supportsA2a":false,"supportsStreaming":false,"inputSchemaRef":null,"outputSchemaRef":null,"dataRegion":null,"contractUpdatedAt":null,"sourceUpdatedAt":null,"freshnessSeconds":null},"invocationGuide":{"preferredApi":{"snapshotUrl":"https://www.xpersona.co/api/v1/agents/clawhub-zeekr0808-hue-multi-model-consensus/snapshot","contractUrl":"https://www.xpersona.co/api/v1/agents/clawhub-zeekr0808-hue-multi-model-consensus/contract","trustUrl":"https://www.xpersona.co/api/v1/agents/clawhub-zeekr0808-hue-multi-model-consensus/trust"},"curlExamples":["curl -s \"https://www.xpersona.co/api/v1/agents/clawhub-zeekr0808-hue-multi-model-consensus/snapshot\"","curl -s \"https://www.xpersona.co/api/v1/agents/clawhub-zeekr0808-hue-multi-model-consensus/contract\"","curl -s \"https://www.xpersona.co/api/v1/agents/clawhub-zeekr0808-hue-multi-model-consensus/trust\""],"jsonRequestTemplate":{"query":"summarize this repo","constraints":{"maxLatencyMs":2000,"protocolPreference":["OPENCLEW"]}},"jsonResponseTemplate":{"ok":true,"result":{"summary":"...","confidence":0.9},"meta":{"source":"CLAWHUB","generatedAt":"2026-10-10T21:55:11.740Z"}},"retryPolicy":{"maxAttempts":3,"backoffMs":[500,1500,3500],"retryableConditions":["HTTP_429","HTTP_503","NETWORK_TIMEOUT"]}},"endpoints":{"dossierUrl":"https://www.xpersona.co/api/v1/agents/clawhub-zeekr0808-hue-multi-model-consensus/dossier","snapshotUrl":"https://www.xpersona.co/api/v1/agents/clawhub-zeekr0808-hue-multi-model-consensus/snapshot","contractUrl":"https://www.xpersona.co/api/v1/agents/clawhub-zeekr0808-hue-multi-model-consensus/contract","trustUrl":"https://www.xpersona.co/api/v1/agents/clawhub-zeekr0808-hue-multi-model-consensus/trust"}},"reliability":{"evidence":{"source":"runtime-metrics","verified":false,"confidence":"low","updatedAt":null,"emptyReason":"No trust, reliability, or runtime telemetry is available."},"trust":{"status":"unavailable","handshakeStatus":"UNKNOWN","verificationFreshnessHours":null,"reputationScore":null,"p95LatencyMs":null,"successRate30d":null,"fallbackRate":null,"attempts30d":null,"trustUpdatedAt":null,"trustConfidence":"unknown","sourceUpdatedAt":null,"freshnessSeconds":null},"decisionGuardrails":{"doNotUseIf":["Contract metadata is missing or unavailable for deterministic execution."],"safeUseWhen":[],"riskFlags":["missing_or_unavailable_contract","trust_data_unavailable","schema_references_missing"],"operationalConfidence":"low"},"executionMetrics":{"observedLatencyMsP50":null,"observedLatencyMsP95":null,"estimatedCostUsd":null,"uptime30d":null,"rateLimitRpm":null,"rateLimitBurst":null,"lastVerifiedAt":null,"verificationSource":null},"runtimeMetrics":{"successRate":null,"avgLatencyMs":null,"avgCostUsd":null,"hallucinationRate":null,"retryRate":null,"disputeRate":null,"p50Latency":null,"p95Latency":null,"lastUpdated":null}},"benchmarks":{"evidence":{"source":"no-benchmark-data","verified":false,"confidence":"low","updatedAt":null,"emptyReason":"No benchmark suites or observed failure patterns are available."},"suites":[],"failurePatterns":[]},"artifacts":{"evidence":{"source":"CLAWHUB","verified":false,"confidence":"high","updatedAt":"2026-10-10T18:44:47.834Z","emptyReason":null},"readme":"Skill: Multi Model Consensus\n\nOwner: zeekr0808-hue\n\nSummary: 多模型决策委员会 — 消除单模型偏见，通过多轮分歧讨论产出客观决策参考。支持3-6个模型同时评审，提供量化投票矩阵和6段式共识报告。触发条件：包含「多模型决策」或「多模型委员会」时自动激活。\n\nTags: latest:1.9.1\n\nVersion history:\n\nv1.8.0 | 2026-05-16T14:47:53.996Z | auto\n\n**Summary:**  \nThis release enforces stricter process control and standardization for the multi-model decision committee, focusing on compliance, template usage, and evaluation transparency.\n\n- Strongly enforces that, upon evaluator/model changes, all available models must be presented for user selection—auto-selection is prohibited.\n- Mandates the use of unified, pre-defined templates (`references/OUTPUT_TEMPLATE.md`) for every process step; organizers must not modify or freestyle output formatting.\n- Introduces and standardizes the required process for model combination recommendations and strict result collection procedures, including mandatory use of `sessions_yield` for subagent result collection.\n- Narrows the consensus report structure from 6 segments to 5, with clearer labeling and requirements.\n- Tightens rules on error handling, round progression, and prohibits advancing to next rounds or final report output before all expected results are received or timeout occurs.\n- Adjusts threshold and evaluation status definitions, listing scoring rules for \"通过\", \"待决策\", and \"有分歧\" with clear icons and consequences.\n\nv1.7.1 | 2026-05-15T03:44:09.578Z | user\n\n环境适配矩阵新增 Feishu direct chat 兜底方案说明\n\nv1.7.0 | 2026-05-07T04:37:34.639Z | user\n\nV1.7.0: 禁止行为升级为4条铁律；异常处理新增框架绕过类型\n\nv1.9.1 | 2026-05-01T00:56:04.986Z | user\n\n删除 version 字段；frontmatter 增加路径约束引用注释，提升审计可见性\n\nv1.6.9 | 2026-05-01T00:53:11.543Z | user\n\n删除 version 字段；frontmatter 增加路径约束引用注释，提升审计可见性\n\nv1.6.8 | 2026-05-01T00:43:50.846Z | user\n\n删除 version 字段；frontmatter 增加路径约束引用注释，提升审计可见性\n\nv1.9.0 | 2026-05-01T00:39:40.393Z | user\n\n删除 version 字段；frontmatter 增加路径约束引用注释，提升审计可见性\n\nv1.6.7 | 2026-04-30T09:32:09.750Z | auto\n\n- 增强「身份纯净原则」：新增评委子Agent工具隔离要求，子Agent仅可通过 `task` 参数评审，不得调用任何工具（如 spawn、Read、Write、Exec、Search 等），所有外部操作统一由组织者管理。\n- 优化第0轮环境兼容性检查流程，明确将其作为系统自动自检以确保运行环境兼容，保障子Agent能正常spawn与回收结果。\n- 相关文档（如 SKILL.md、README.md 等）同步修订，反映治理细则和流程说明的新要求。\n- 其余决策流程、判定规则与核心报告结构维持原有设计。\n\nv1.6.6 | 2026-04-30T09:14:30.118Z | auto\n\nVersion 1.6.6\n\n- 明确限制多模型评审委员数量为3–6个（原为3–13个）。\n- 新增安全白名单说明：`Read`工具和`Write`工具的可访问路径范围进一步细化和限制，禁止读取用户敏感目录和主目录隐藏文件。\n- 删除 references/VERIFICATION_CASE.md 等部分不相关或无用的参考内容。\n- 文档内容优化与结构细化，对技术约束和工具权限范围进行详细补充。\n- 其他细微表述调整，提升说明准确性与用户操作安全性。\n\nv1.6.5 | 2026-04-30T03:58:50.614Z | user\n\nSecurity hardening: remove Exec from allowed-tools, reduce max concurrent sub-agents 13->6, add Write tool path restriction, model-routing early termination rule in TROUBLESHOOTING.md\n\nv1.6.4 | 2026-04-29T12:21:21.610Z | auto\n\nmulti-model-consensus v1.6.4\n\n- 更新版本号至 1.6.4。\n- 移除 references/EXECUTION_EXAMPLES.md 参考文档条目与文件，无其他文档更新。\n- 其余功能、流程与参数保持不变。\n\nv1.6.3 | 2026-04-29T12:03:42.849Z | auto\n\n**重大更新：引入决策点机制，优化流程与通知，全面提升多模型决策闭环和适配性。**\n\n- 决策点级别评审：每轮按决策点拆分评分和判定，支持单方案/二选一/多决策点复合任务。\n- 弃用持久化状态文件，改为会话上下文变量进行轻量级实时状态跟踪。\n- 统一并推送式（Push-based）子Agent结果回收机制，详细规范跨平台调用和回流方法。\n- 环境适配细化，webchat等环境每轮后即时展示评分矩阵和判定结果；thread 类环境维持有序消息。\n- 配置参数和判定方式精简及明确，收敛分差阈值参数废弃，所有评审基于通过阈值与决策点状态。\n- 新增详细 TROUBLESHOOTING.md 文档，系统化汇总异常类型与处理标准。\n\nv1.5.6 | 2026-04-26T09:47:23.079Z | user\n\nV1.5.6: Judges prohibited from spawning sub-agents; organizer controls all task distribution\n\nv1.5.5 | 2026-04-26T08:36:46.221Z | user\n\nSync full USER_GUIDE content to README\n\nv1.5.4 | 2026-04-26T08:29:53.676Z | user\n\nAdd docs/README.md for ClawHub display page\n\nv1.5.3 | 2026-04-26T08:12:24.508Z | user\n\nAdd contact email to USER_GUIDE.md footer\n\nv1.5.2 | 2026-04-26T08:00:30.572Z | auto\n\nVersion 1.5.2 summary: Adds configurable consensus verdict method; updates documentation and schema references.\n\n- Introduced configurable \"verdict rule\" (判定方式) with support for 全票通过 (unanimous), 均分通过 (average), and 多数票通过 (majority) threshold modes; unanimous remains the default.\n- Updated documentation across SKILL.md, USER_GUIDE.md, and related references to clarify the new configurable judgment modes and workflow.\n- Added `references/SCHEMA.md` for standardized state field definitions; removed previous schema version file.\n- Enhanced the execution rules, sample flows, and parameter controls for transparency and clarity.\n- Synchronized all version metadata to 1.5.2.\n\nv1.2.5 | 2026-04-25T05:56:53.173Z | user\n\nV1.2.5 (2026-04-25):\n- ENFORCEMENT: Add mandatory threshold check rules (强制判定规则) in SKILL.md\n  - Round 1: MUST check average score vs threshold before report\n  - Round 2: MUST check convergence before skipping Round 3\n  - Add forbidden items for skipping threshold/convergence checks\n  - State machine now has explicit checkpoint enforcement\n\nV1.2.4 (2026-04-25): Remove anonymous委员 from prompt templates\nV1.2.3 (2026-04-25): First-time user flow clarification\nV1.2.2 (2026-04-25): SYNC after hotfix\nV1.2.1 (2026-04-25): Remove hardcoded model list\n\nv1.2.4 | 2026-04-25T04:46:18.606Z | user\n\nV1.2.4 (2026-04-25):\n- CRITICAL FIX: OUTPUT_TEMPLATE.md - remove \"匿名委员\" from all 3 round prompt templates; replace with real-name format \"{模型名称}（实名委员）\"\n\nV1.2.3 (2026-04-25): Clarify first-time user flow\nV1.2.2 (2026-04-25): SYNC after v1.2.1 hotfix\nV1.2.1 (2026-04-25): Remove hardcoded model list; dynamic scan\n\nv1.2.3 | 2026-04-25T04:25:33.317Z | user\n\nV1.2.3 (2026-04-25):\n- UX FIX: Clarify first-time user flow in SKILL.md and Quick Start guide:\n  1. Auto-scan local openclaw.json models.providers\n  2. Default to first 3 models as committee\n  3. Prompt user to confirm or modify before starting decision\n\nV1.2.2 (2026-04-25): SYNC after v1.2.1 hotfix\nV1.2.1 (2026-04-25): Remove hardcoded model list; dynamic scan\n\nv1.2.2 | 2026-04-25T04:13:38.407Z | user\n\nV1.2.2 (2026-04-25):\n- SYNC: Full file sync to GitHub and ClawHub after v1.2.1 hotfix\n- FIX: Remove hardcoded model list from SKILL.md; dynamic scan of openclaw.json models.providers\n- REFACTOR: EXECUTION_EXAMPLES.md - removed duplicate prompt templates, consolidated to OUTPUT_TEMPLATE.md\n- REFACTOR: VERIFICATION_CASE.md - removed duplicate workflow steps, keep only verification checklist\n- DOCS: USER_GUIDE.md - added V1.2.0 and V1.2.1 changelog entries, bilingual alignment\n\nV1.2.1 (2026-04-25): Already published (same content as 1.2.2)\nV1.2.0 (2026-04-25): Restored public ClawHub version\n\nv1.2.1 | 2026-04-25T04:02:32.594Z | user\n\nv1.2.1: Fix hardcoded model list - now dynamically scans local openclaw.json for available models; removes fixed A1/A2/A4... table and replaces with real-time model detection\n\nv1.2.0 | 2026-04-24T06:22:52.004Z | user\n\nv1.2.0: clean package — only essential files (SKILL.md + 5 refs); docs/USER_GUIDE.md bilingual for global community; CHANGELOG/PROPOSAL removed from package\n\nv1.1.6 | 2026-04-24T06:14:24.692Z | user\n\nv1.1.6: model list now dynamic — auto-detect available models on first launch, default to first 3\n\nv1.1.5 | 2026-04-24T04:59:17.057Z | user\n\nv1.1.5: remove backup/ directory from package; bilingual description in frontmatter; ClawHub README now EN/CN bilingual\n\nv1.1.4 | 2026-04-24T04:48:23.709Z | user\n\nv1.1.4: SKILL.md and references/ converted to pure Chinese (token optimized); docs/ retains full bilingual USER_GUIDE.md for global community\n\nv1.1.3 | 2026-04-24T04:30:32.351Z | user\n\nv1.1.3: SKILL.md converted to bilingual EN/CN format; docs/ duplicate files removed; PROPOSAL.md moved to root\n\nv1.1.2 | 2026-04-24T04:08:50.161Z | user\n\nv1.1.2: rename GUIDE.md to USER_GUIDE.md; all references docs bilingual format; full repository restructure\n\nv1.1.1 | 2026-04-24T03:24:11.877Z | user\n\nv1.1.1: restructure GUIDE.md - English guide on top, Chinese below, separator line; delete duplicate user manual; translation optimizations\n\nv1.1.0 | 2026-04-24T02:53:41.206Z | user\n\nInitial release v1.1.0\n\nv1.0.0 | 2026-04-24T02:53:05.872Z | user\n\nInitial release v1.0.0\n\nArchive index:\n\nArchive v1.8.0: 10 files, 55400 bytes\n\nFiles: _meta.json (140b), docs/README.md (23803b), docs/USER_GUIDE.md (38739b), README.md (23984b), references/OUTPUT_TEMPLATE.md (15873b), references/SCHEMA.md (1314b), references/STATE_MACHINE.md (3054b), references/TROUBLESHOOTING.md (5964b), references/VERIFICATION_CASE.md (4333b), SKILL.md (20664b)\n\nFile v1.8.0:SKILL.md\n\n---\nname: multi-model-consensus\ndescription: 多模型决策委员会 — 消除单模型偏见，通过多轮分歧讨论产出客观决策参考。支持3-6个模型同时评审，提供量化投票矩阵和5段式共识报告。触发条件：包含「多模型决策」或「多模型委员会」时自动激活。\nallowed-tools: SessionsSpawn,SessionsSend,SessionsHistory,Read,Write\nmetadata:\n  openclaw:\n    emoji: \"🏛️\"\n    requires:\n      capability: sessions_spawn\n    # 路径约束：见正文\"技术约束\"章节\n---\n\n# 🏛️ 多模型决策委员会\n\n专为 OpenClaw 设计的\"数字智库\"，通过多模型独立评审与共识合成，消除单模型偏见，为复杂决策提供客观参考。\n\n---\n\n\n## 🎯 插件定位\n\n**多模型决策委员会** 是专为 OpenClaw 设计的\"数字智库\"。它支持调动多个不同架构的大模型（如 GPT、Gemini、豆包等）同时对同一任务进行协同思考、对抗辩论与共识合成，有效消除单模型偏见，综合性给出最优建议。\n\n---\n\n## 核心优势\n\n- **去中心化决策**：交叉验证不同模型逻辑，确保方案的严谨性。\n- **对抗性评审**：通过模型间的互评，快速锁定潜在的逻辑漏洞。\n- **弹性扩容**：根据任务难度，随时增加或减少\"决策委员\"的数量。\n- **消除偏见**：多模型独立评审，避免单一AI的认知盲区。\n- **量化决策**：6维度评分 + 加权矩阵，结论有据可查。\n- **透明可信**：实名委员制、模型身份透明。\n- **灵活配置**：参数可调，如决策成员人数、决策轮次、通过阈值等，均可可自定义。\n\n---\n\n## 核心治理原则：身份纯净\n\n为了确保结果的绝对客观，本插件强制执行 **\"身份纯净原则\"**：\n- **角色透明化**：决策委员会成员均会注明其身份（大模型名称），保持可视化透明。\n- **严禁设定角色**：禁止给委员设定诸如\"架构师\"、\"审计员\"等身份标签。\n- **严禁引导提示**：任务下发时不得针对评审委员设置诱导性提示词或诱导性身份角色，防止产生的偏见。\n- **评委工具隔离**：评委子Agent仅通过 `task` 参数接收评审内容，独立输出结论；不得调用任何工具（包括 spawn、Read、Write、Exec、Search 等），所有外部操作由组织者统一管理\n- **评审不重复原则**：若组织者（即当前使用模型）是评审委员会成员，则组织者提交内容即视同为其在本轮的评审结果，无需重复自评。若组织者不是评审委员会成员，则仅负责组织实施和汇总评委意见，不参与评审。\n\n---\n\n## 使用方法\n\n- **启动**：在OpenClaw中，当用户输入的内容包含「多模型决策」或「多模型委员会」时，自动激活多模型决策委员会。首次使用时会提醒用户进入配置模式，并选择模型组合。\n- **配置**：根据需要，可配置参数，如：轮数、阈值、委员数量、委员模型等。\n\n### 参数配置\n\n配置参数可采用指令方式，如：\"修改配置\"、\"调整参数\"、\"增加委员\"、\"确认配置\"。\n\n| 指令关键词 | 可调参数 | 说明 |\n|:---|:---|:---|\n| \"换模型\" | 委员模型组合 | 从用户本地接入的模型列表中选择模型组合加入。用户要求调整评委时，组织者**必须**读取 `openclaw.json` 列出当前环境下所有可接入的大模型名单（含Provider、上下文、最大输出等关键信息），供用户选择。 |\n| \"改轮数\" | 决策轮数 | 默认3轮，日常事务可设为2轮，重大决策可设3-6轮 |\n| ~~\"收敛分差阈值\"~~ | ~~收敛判定分差~~ | ~~已废弃~~ |\n| \"改阈值\" | 判定阈值 | **通过阈值**：默认≥90%；**待定阈值**：默认[60%, 90%)；**否定阈值**：默认<60% |\n| \"改判定方式\" | 通过判定方式 | **全票通过** |\n| \"增加委员\" / \"减少委员\" | 委员数量 | 支持2-6个模型同时评审。用户要求增减委员时，组织者**必须**先列出当前环境下所有可接入的大模型名单，供用户指定。 |\n| \"确认配置\" | 确认当前配置 | 确认当前参数后开始决策 |\n\n### 评委调整规则（铁律）\n\n当用户表达\"换评委\"、\"换模型\"、\"调整评委\"、\"增减委员\"等意图时，组织者**必须**执行以下步骤：\n\n1. **列出全部可用模型**：读取 `openclaw.json` 中所有 provider 的 models 列表，完整展示：\n   - 模型ID（如 `gpt-5.5`、`gemini-3.1-pro-preview`）\n   - 模型名称（如 GPT-5.5、Gemini 3.1 Pro）\n   - Provider来源（如 n1n、n1n-gemini、ark-cn）\n   - 上下文窗口（contextWindow）\n   - 最大输出（maxTokens）\n2. **推荐组合**：基于任务特点，给出1-3组推荐搭配（如\"逻辑+中文\"、\"深度+性价比\"）。\n3. **等待用户指定**：禁止直接替用户选择，必须等待用户明确指定模型组合后方可继续。\n\n### 参数建议\n\n- **模型选择**：建议至少包含 1 个具备强逻辑推理能力的模型和 1 个具备强中文语境理解能力的模型。\n- **轮次建议**：日常事务建议 2 轮；涉及架构，资金、核心规则的决策，建议 3-5 轮。\n- **阈值设定**：\n  - **严谨型**：≥95%（全票通过）\n  - **效率型**：≥75%（多数通过）\n\n---\n\n## 决策流程（3轮决策机制）\n\n多模型决策委员会采用 3 轮决策机制，基于决策点级别评审，每轮进行100分制评分，最终给出决策结果。\n\n---\n\n### 第 0 轮：准备与框架确认 (Preparation & Framework Confirmation)\n\n**决策准备**：组织者（当前使用模型）接收到决策任务后，将用户背景、需求、待审方案汇总，拆解为决策点及权重，发起决策申请。\n\n> 📌 **标准模板**：组织者须完整套用 `references/OUTPUT_TEMPLATE.md` 中的 **T1 决策准备确认模板**，不得自行增删字段或自由发挥。\n\n---\n\n### 运行时自检（第 0 轮环境兼容性检查）\n\n组织者在启动评委前，执行环境自检以确保子Agent可正常 spawn 和返回结果。此步骤为系统自动执行，确保运行环境兼容性。\n\n**自检方法**：通过 `session_status` 或检查当前会话元数据，确认通道类型。\n\n**环境适配矩阵**：\n\n| 环境 | Runtime | Mode | 额外参数 | 结果回收方式 |\n|:---|:---|:---|:---|:---|\n| **Webchat** | `subagent` | `\"run\"` | 无 | `subagent_announce` 事件回流 |\n| **Feishu / Telegram / Discord** | `acp` | `\"session\"` | `streamTo: \"parent\"`, `thread: true` | 流式回流至 organizer 会话 |\n| **Unknown/不确定** | `subagent` | `\"run\"` | 无 | `subagent_announce` 事件回流 + **主动查询兜底** |\n\n---\n\n### 第 1 轮：独立评估 (Blind Evaluation)\n\n**动作说明**：组织者接收到决策确认后，将准备流程相关内容分发至各选定模型的独立子会话中。每位委员在**完全无法感知他方意见**的情况下进行背对背评审。\n\n**评审维度**：组织者将方案拆解为若干「决策点」，每个决策点对应一个具体评审维度。评委对每个决策点逐项打分（0-100分制）。整体方案得分为各决策点评分的加权平均值，由组织者自动计算，不再由评委单独打分。\n\n**评审维度示例**：方案逻辑合理性、完整性、潜在风险分析、实施资源消耗预估、可行性评审、优化建议。\n\n**评审时长**：不超过120秒/轮。若120秒仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：组织者汇总各委员对所有决策点的评分结果，标记出不通过（评分 < 阈值）的决策点，提取为「不通过决策点清单」。\n\n---\n\n### 第 2 轮：分歧讨论（Dispute Discussion）\n\n**动作说明**：组织者将第1轮中标记为「未通过」的决策点（🟡待定项+🔴不通过项）整理为「未通过决策点清单」，由组织者综合评委建议，形成调整优化方向后，将这些未通过项及调整方向一并下发给所有委员，要求评委**只针对这些未通过决策点**进行讨论并重新评分。已通过的决策点不再讨论。\n\n**评审时长**：不超过120秒/轮。若120秒仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：汇总本轮重新评分结果，**若所有决策点全员通过则直接输出最终报告；若仍有未通过项，由组织者综合评委建议形成调整优化方向后，进入下一轮**。\n\n---\n\n### 最后一轮：最终辩论 (Final Duel)\n\n**特殊说明**：无论配置多少轮，最终辩论始终是**最后一轮**。\n- 3轮配置：第3轮为最终辩论\n- 5轮配置：第5轮为最终辩论\n- N轮配置：第N轮为最终辩论\n\n**动作说明**：经过分歧讨论后，仍有未通过项时启用。组织者将仍未通过的决策点（🟡待定项+🔴不通过项）**由组织者综合评委建议，形成调整优化方向后**，再次下发给所有委员，要求评委进行最后一轮讨论和重新评分。\n\n**评审时长**：不超过120秒/轮。若120秒仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：若所有决策点全员通过，输出最终报告；若仍有未通过项，**由组织者综合评委建议形成调整优化方向后**，将未通过项标注状态后输出最终报告。\n\n---\n\n### 轮次扩展说明\n\n| 扩展类型 | 说明 |\n|:---|:---|\n| 增加轮次（4轮及以上） | 倒数第2轮参照「分歧讨论」；最后一轮参照「最终辩论」；中间轮次循环分歧讨论 |\n| 减少轮次（3轮以下） | 第1轮参照「独立评估」；第2轮直接输出最终报告（视为最后一轮） |\n\n---\n\n## 自动化通知规则\n\n为防止评委评审时间过长导致用户对当前状态不知情，组织者必须在每轮评审开始前主动通知用户：\n\n- **进入第1轮**：通知用户「各位评委已开始独立评估，请稍候」\n- **进入第2轮**：通知用户「各位评委已开始分歧讨论，请稍候」\n- **进入最后一轮**：通知用户「各位评委已开始最终辩论，请稍候」\n- **输出最终报告**：通知用户「最终报告已生成」\n\n---\n\n## 中间结果输出规则（webchat 适配）\n\n当环境为 webchat 时，用户无法像 Feishu thread 那样看到有序的中间消息。为提升用户体验，**每轮结束后应立即向用户展示本轮汇总**，而不是等到全部轮次结束后再统一输出。\n\n| 环境 | 中间结果处理 |\n|:---|:---|\n| **webchat** | 每轮结束后立即输出本轮汇总矩阵 + 判定结果 + 提示语，然后继续下一轮 |\n| **Feishu / Discord（thread 模式）** | 按原流程执行，threaded 消息天然有序，无需特殊处理 |\n| **Unknown / 不确定** | 默认视为 webchat，执行中间输出规则 |\n\n**webchat 环境每轮输出模板**：\n\n> 📌 组织者须严格套用 `references/OUTPUT_TEMPLATE.md` 中的对应模板：\n> - 第 1 轮汇总 → **T3 第1轮汇总模板**\n> - 第 2 轮汇总 → **T5 第2轮汇总模板**\n> - 最终报告 → **T7 最终报告模板**\n> \n> 禁止自行增删字段或自由发挥格式。\n\n---\n\n## 通过判定规则\n\n每轮结束后**必须**进行判定，不可跳过。\n\n**【核心概念：决策点】**\n每个方案的评审内容拆解为若干「决策点」，评委对每个决策点评分（0-100分）。通过判定基于决策点级别进行。\n\n**【判定方式】（默认全票通过制）**\n- **全票通过**（默认）：所有评委在所有决策点上的评分均需≥阈值，任一评委在任一决策点低于阈值即判定为不通过\n\n**通过判定逻辑**：\n\n| 轮次 | 判定时机 | 判定结果 | 后续动作 |\n|:---|:---|:---|:---|\n| 第1轮结束后 | 汇总所有决策点评分 | **全员通过**（所有决策点所有评委均≥阈值） | 直接输出最终报告 |\n| 第1轮结束后 | 汇总所有决策点评分 | **非全员通过**（存在未通过决策点） | 提取未通过的决策点，进入第2轮 |\n| 第2轮结束后 | 汇总不通过决策点 | **全员通过** | 直接输出最终报告 |\n| 第2轮结束后 | 汇总未通过决策点 | **仍有未通过** | 进入最后一轮最终辩论 |\n| 最后一轮结束后 | 汇总不通过决策点 | **全员通过** | 直接输出最终报告 |\n| 最后一轮结束后 | 汇总未通过决策点 | **仍有未通过** | 输出最终报告，标注各决策点状态 |\n| 任意轮超限 | — | — | 标记超时评委，以已返回结果继续汇总 |\n\n**各轮次讨论范围**：\n\n| 轮次 | 讨论范围 |\n|:---|:---|\n| 第2轮（分歧讨论） | 仅讨论上一轮中未通过的决策点（🟡待定项+🔴不通过项） |\n| 第3轮（最终辩论） | 仅讨论第2轮中仍未通过的决策点（🟡待定项+🔴不通过项） |\n| 后续轮次 | 循环「分歧讨论」逻辑，直至所有决策点全员通过或已达最后一轮 |\n\n---\n\n## 最终输出：5段式共识报告\n\n决策完成后，输出标准化报告，包含：\n\n1. **投票记录** — 量化评分矩阵（含决策点级别评分）\n2. **汇总说明** — 各委员核心论点（注明大模型）\n4. **决策点通过清单** — 各决策点通过状态\n6. **风险提示** — 主要风险与缓解建议\n\n**决策点通过状态标注**：\n\n| 状态 | 含义 | 标注 |\n|:---|:---|:---|\n| 全员通过 | 所有评委在该决策点评分均≥阈值 | ✅ 绿色 — 通过 |\n| 待决策 | 推荐顺序一致但有评委评分低于阈值 | 🟡 黄色 — 待用户确认 |\n| 有分歧 | 推荐顺序不一致 | 🔴 红色 — 有分歧，附原因 |\n\n---\n\n## 子Agent结果回收机制（重要）\n\n### 结果回收机制（按环境区分）\n\n#### Webchat / 标准环境 — Push-based\n\n组织者使用 **Push-based（推送式）** 回收子Agent结果：\n\n```\nSpawn 子Agent → sessions_yield 挂起 → 收到 subagent_announce 事件 → 提取结果 → 更新跟踪\n```\n\n**关键步骤**：\n\n1. **Spawn 子Agent** 后，立即调用 `sessions_yield` 主动结束当前 turn\n2. **等待** OpenClaw 以 inter-session message 形式推送 `subagent_announce` 事件\n3. **识别** 消息中的 `BEGIN_UNTRUSTED_CHILD_RESULT` 块，提取评审结果\n4. **记录** 结果到会话上下文\n5. **检查** 是否所有评委都已返回\n\n⚠️ **禁止行为（铁律）**：\n\n1. **严禁跳过 sessions_yield**：spawn 评委后必须立即调用 `sessions_yield` 挂起，等待 `subagent_announce` 事件触发结果回流。严禁在结果未到达前主动调用 `sessions_history` 查询结果。\n\n2. **严禁跳轮宣布**：宣布进入下一轮前，必须先完成本轮所有评委结果的收集。未收集完全就宣布进入下一轮等同于「规则执行偏差」。\n\n3. **主动查询兜底机制（Unknown 环境适用）**：当 spawn 后 120 秒内未通过 `subagent_announce` 收到全部评委结果时， organizer 应主动调用 `subagents list` 查询子Agent状态。若子Agent状态为 `done` 但结果未回流， organizer 应调用 `sessions_history` 提取结果。此兜底机制仅限 push-based 失效时使用，非默认行为。\n\n4. **严禁在结果未全部到达时输出最终报告**：必须等待所有评委返回，或等待超时后以已到达结果继续，同时标记未到达评委为「超时-未提交」。\n\n### 标准调用示例\n\n**Webchat 环境（推荐默认）**：\n```yaml\nSessionsSpawn:\n  runtime: \"subagent\"\n  mode: \"run\"\n  model: \"n1n/gpt-5.4\"\n  task: \"评审任务...\"\n  timeoutSeconds: 180\n```\n\n**Feishu/Discord 环境**（支持 thread 的通道）：\n```yaml\nSessionsSpawn:\n  runtime: \"acp\"\n  mode: \"session\"\n  thread: true\n  streamTo: \"parent\"\n  model: \"n1n/gpt-5.4\"\n  task: \"评审任务...\"\n  timeoutSeconds: 180\n```\n\n### 超时处理\n\n单轮限时 120 秒（2分钟）。若超时：\n- 标记该评委为「超时-未提交」\n- 以已返回的评委结果继续汇总\n- 最终报告中注明超时评委\n\n---\n\n## 轻量级状态跟踪（替代状态文件）\n\n组织者使用**会话上下文变量**跟踪评审状态，无需维护持久化状态文件。\n\n**状态跟踪格式**（每轮开始时初始化）：\n```\n[本轮状态跟踪]\nRound: {1/2/3}\nSpawned: {N}\nReceived: {M} ({评委A} ✅, {评委B} ✅)\nPending: {N-M} ({评委C} ⏳)\n```\n\n每收到一个子Agent结果，更新此跟踪块。全部完成后清空并进入下一轮。\n\n---\n\n## 异常处理规则\n\n以下3类异常的处理标准：\n\n| 异常类型 | 处理标准 |\n|:---|:---|\n| 状态同步错误 | 组织者检测到会话上下文中的状态跟踪与实际回收结果不符时，应立即中止当前流程，通知用户并提示重新发起决策 |\n| 规则执行偏差 | 组织者在执行中发现某委员的评审结果违反本skill规定的流程（如跳轮、跳过通过判定等），应要求该委员重新按规则执行，不得擅自修改委员结论 |\n| 超时处理 | 单轮超过120秒仍有评委未返回结果时，组织者对超时评委标记「超时-未提交」，以已返回的评委结果进行汇总和判定 |\n| 结果未回流 | 若 spawn 后未收到 subagent_announce 事件，检查是否使用了 sessions_yield 挂起，或 runtime/mode 配置是否正确 |\n| 框架绕过 | 组织者绕过 skill 机制私自操作（如绕 sessions_yield 直接查 history、跳轮宣布等），应立即承认并通知用户，由用户决定是重新发起还是继续 |\n\n---\n\n## 技术约束\n\n### 工具权限范围（安全白名单）\n- `Read` 工具仅限读取以下路径：\n  - skill 目录内文件（SKILL.md、references/、docs/ 等）\n  - `~/.openclaw/workspace/tmp_mmc_*.md`（临时报告文件）\n  - 严禁读取用户主目录下的隐藏文件（`.ssh/`、`.env/`、`.aws/` 等）\n- `Write` 工具仅限写入以下路径：\n  - `~/.openclaw/workspace/memory/` 目录（报告归档）\n  - `~/.openclaw/workspace/tmp_mmc_*.md`（临时报告文件）\n  - `~/Desktop/*`（最终报告输出到桌面）\n  - 禁止写入其他任何路径\n\n- 每个子会话最大执行时间：**120秒**（可通过 `timeoutSeconds` 配置，最大 300 秒）\n- 最大并发子会话数：**6个**（与委员数量上限一致）\n- ~~状态文件有效期：**24小时**~~（V1.6.0 起不再使用状态文件，改用会话上下文跟踪）\n- 报告自动归档路径：`~/.openclaw/workspace/memory/MONTHLY/mmd_<date>.md`\n\n---\n\n## 参考文档\n\n| 文档 | 说明 |\n|:---|:---|\n| [references/STATE_MACHINE.md](references/STATE_MACHINE.md) | 状态流转规则 + 异常处理（概念参考，V1.6.0 起实际使用会话上下文跟踪） |\n| [references/OUTPUT_TEMPLATE.md](references/OUTPUT_TEMPLATE.md) | **T1/T3/T5/T7 固化模板** + 评委分发 Prompt（T2/T4/T6） |\n| [references/SCHEMA.md](references/SCHEMA.md) | 状态文件字段规范（已归档，V1.6.0 起不再强制使用） |\n| [references/TROUBLESHOOTING.md](references/TROUBLESHOOTING.md) | 常见失败模式与排查指南（V1.6.0 新增） |\n\n---\n\n## 更新日志\n\n| 版本 | 日期 | 变更内容 |\n|:---:|:---:|:---|\n| V1.8.0 | 2026-05-16 | **重大更新**：模板体系重构——T1/T3/T5/T7 全面修订；6段式报告改为5段式；判定方式固定为全票通过；评审方案固定为决策点拆分评审；进入下一轮的决策点须由组织者综合评委建议形成调整优化方向后分发；T4/T6 评委Prompt删除二选一，统一为态度追踪+重新评分；SKILL.md规则与模板描述完全对齐 |\n| V1.7.4 | 2026-05-16 | 环境适配矩阵更新：Feishu direct chat 改用 `acp` + `session` + `thread: true` + `streamTo: \"parent\"`，结果流式回流至 organizer；移除旧的飞书 direct chat 特殊处理段落和 sessions_history 兜底例外 |\n| V1.7.3 | 2026-05-16 | 固化 T1/T3/T5/T7 四阶段模板：T1 决策准备确认模板、T3 第1轮汇总模板、T5 第2轮汇总模板、T7 最终报告模板；T2/T4/T6 确认为评委分发 Prompt，组织者严格按模板执行 |\n| V1.7.2 | 2026-05-15 | 新增「评委调整规则（铁律）」：用户要求调整评委时，必须列出当前环境下所有可接入的大模型名单，禁止直接替用户选择 |\n| V1.7.1 | 2026-05-15 | 环境适配矩阵新增 Feishu direct chat 兜底方案说明 |\n| V1.7.0 | 2026-05-07 | 禁止行为升级为4条铁律（严禁跳过sessions_yield/严禁跳轮宣布/严禁轮询/严禁提前输出）；异常处理规则新增「框架绕过」类型 |\n| V1.6.0 | — | 改用会话上下文跟踪，废除状态文件 |\n| V1.5.0 | — | webchat环境适配 |\n\n---\n\n🏛️ **兼听则明，万模共鉴。**\n\nFile v1.8.0:docs/README.md\n\n# Multi-Model Consensus Council — Operation & User Guide\n\n## 📋 Version History\n\n| Version | Date | Changes |\n|:---|:---|:---|\n| V1.0.0 | 2026-04-19 | Initial release: core framework, 3-round convergence, 6-section report |\n| V1.1.0 | 2026-04-24 | Added English operation guide (merged with Chinese, English first); translation reviewed and approved by multi-model committee (Gemini/Doubao/GLM, avg 81/100); wording optimizations |\n| V1.1.3 | 2026-04-24 | SKILL.md and references/ converted to pure Chinese for AI readability; docs/ retains full bilingual documentation for global community |\n| V1.2.0 | 2026-04-25 | Restored public ClawHub version (was local customized); sync all files |\n| V1.2.1 | 2026-04-25 | **CRITICAL FIX**: Removed hardcoded model list (A1/A2/A4...) from SKILL.md; replaced with dynamic scan via `openclaw models list`; examples updated to use generic `[Model A/B/C]` placeholders |\n| V1.2.2 | 2026-04-25 | SYNC: Full file sync to GitHub and ClawHub after v1.2.1 hotfix |\n| V1.2.3 | 2026-04-25 | **UX FIX**: Clarify first-time user flow: auto-scan models → default to first 3 → prompt user to confirm/modify before starting decision |\n| V1.2.4 | 2026-04-25 | **CRITICAL FIX**: OUTPUT_TEMPLATE.md - remove \"匿名委员\" from all 3 round prompt templates; replaced with real-name format `{模型名称}（实名委员）` |\n| V1.2.5 | 2026-04-25 | **ENFORCEMENT**: SKILL.md - add mandatory threshold check rules; add forbidden items for skipping threshold/convergence checks; state machine now has explicit checkpoint enforcement |\n| V1.2.6 | 2026-04-25 | **Real-name Transparency**: Judges use standard model names; 3-layer architecture (organizer/judge/sub-agent); 100-point scoring; Round 0 preparation phase added |\n| V1.5.0 | 2026-04-25 | Sync all document versions to V1.5.0; renamed SCHEMA.md; full English content aligned with Chinese |\n| V1.5.1 | 2026-04-26 | Added round-start user notification rule; added convergence score threshold parameter; added exception handling rules section |\n| V1.5.2 | 2026-04-26 | Added judgment method configurable parameter; clarified unanimous-pass rule as default |\n\n---\n\n## 🎯 Plugin Overview\n\n**Multi-Model Consensus Council** is a \"Digital Think Tank\" designed for OpenClaw. It coordinates multiple large language models (e.g., GPT, Gemini, Doubao) to collaboratively analyze, debate, and synthesize decisions, effectively eliminating single-model bias.\n\n### Core Advantages\n\n- **Decentralized Decision-Making**: Cross-validates reasoning across models, ensuring rigor.\n- **Adversarial Review**: Rapidly identifies logical flaws through peer evaluation.\n- **Elastic Scaling**: Adjust the number of \"decision committee members\" anytime based on task complexity.\n- **Bias Elimination**: Multi-model independent review avoids single-AI cognitive blind spots.\n- **Quantitative Decision-Making**: 6-dimension scoring + weighted matrix, conclusions are evidence-based.\n- **Transparency & Trust**: Real-name committee system, model identity visible.\n- **Flexible Configuration**: All parameters adjustable — number of members, decision rounds, pass threshold, etc.\n\n---\n\n## ⚖️ Core Governance Principle: Identity Purity\n\nTo ensure absolute objectivity, this plugin enforces the **\"Identity Purity Principle\"**:\n\n- **Identity Transparency**: All committee members are identified by their model name (e.g., Kimi K2.6, GPT-5.4), maintaining visual transparency.\n- **No Role Assignment**: Committee members are forbidden from having preset roles such as \"architect\" or \"auditor\".\n- **No Biased Prompting**: Tasks must not include inducing prompts or identity roles targeted at committee members, preventing bias from prompt engineering.\n- **Conclusion Convergence Principle**: If a committee member calls a sub-agent based on their own judgment, they must converge the sub-agent's output within the same round before submitting their final conclusion to the organizer.\n- **No Duplicate Review Principle**: If the organizer (the current model in use) is a committee member, their submitted content counts as their Round 1 review result — no duplicate self-review is required. If the organizer is not a committee member, they are responsible only for organizing and summarizing; they do not participate in the review.\n\n---\n\n## How to Use\n\n- **Activation**: In OpenClaw, when the user's input contains 「多模型决策」 or 「多模型委员会」, the Multi-Model Consensus Council activates automatically. On first use, the system will prompt you to enter configuration mode and select the model lineup.\n- **Configuration**: Adjust parameters as needed, such as rounds, threshold, number of members, and member models.\n\n### Parameter Configuration\n\n| Command Keyword | Adjustable Parameter | Description |\n|:---|:---|:---|\n| \"换模型\" / \"Change models\" | Committee model mix | Select models from the locally connected model list |\n| \"改轮数\" / \"Change rounds\" | Decision rounds | Default 3 rounds; 2 for routine tasks; 3-6 for critical decisions |\n| \"改阈值\" / \"Change threshold\" | Judgment thresholds | **Pass threshold**: default ≥90%; **Pending threshold**: default [72%, 90%); **Reject threshold**: default <72% |\n| \"改判定方式\" / \"Change judgment method\" | Judgment method | **Unanimous-pass (default)**: all judges must meet threshold; Average-pass: average score meets threshold; Majority-pass: more than half meet threshold |\n| ~~\"收敛分差阈值\"~~ / ~~\"Convergence threshold\"~~ | ~~Score convergence threshold~~ | ~~Deprecated — replaced by decision-point pass logic~~ |\n| \"增加委员\" / \"Add member\" | Committee size | Increase number of committee members |\n| \"减少委员\" / \"Remove member\" | Committee size | Decrease number of committee members |\n| \"确认配置\" / \"Confirm config\" | Confirm current config | Confirm current parameters and begin decision |\n\n### Parameter Recommendations\n\n- **Model Selection**: Recommend at least 1 model with strong logical reasoning and 1 model with strong Chinese-language comprehension.\n- **Round Recommendations**: Routine tasks: 2 rounds; Decisions involving architecture, funding, or core rules: 3-5 rounds.\n- **Threshold Settings**:\n  - **Strict**: ≥95% (near-unanimous)\n  - **Efficient**: ≥75% (majority)\n\n---\n\n## Decision Process (3-Round Convergence Mechanism)\n\nThe Multi-Model Consensus Council adopts a 3-round convergence mechanism, with each round using a **100-point scoring system**, ultimately producing a decision result.\n\n---\n\n### Round 0: Preparation (Preparation)\n\n**Task**: Upon receiving a decision task, the **organizer** (the current model in use) summarizes the user background, requirements, and the proposal to initiate the decision application. Content format:\n\n- **Project Name**: {Project Name}\n- **Project Background**: {Project Background}\n- **Project Requirements**: {Project Requirements}\n- **Proposal Under Review**: {Proposal Under Review}\n- **Decision Members**: {Decision Members: pull from configured model lineup}\n- **Decision Rounds**: {Decision Rounds: pull from configured round count}\n- **Threshold Settings**: {Threshold Settings: pull from configured threshold}\n- **Submit Decision**: {Submit whether the user wants to begin decision}\n\n---\n\n### Round 1: Independent Evaluation (Blind Evaluation)\n\n**Action**: Upon receiving the decision confirmation, the organizer distributes the preparation content to each selected model's independent sub-session. Each committee member conducts closed-door review **without any awareness of other members' opinions**.\n\n**Review Dimensions**: Logical completeness, risk analysis, resource cost estimation, feasibility review, optimization suggestions — all scored on a **100-point scale**.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: The organizer summarizes each committee member's initial review report, voting results, scores, and disagreement points or controversy areas.\n\n---\n\n### Convergence Rounds (All rounds except the final round)\n\n**Action**: The organizer summarizes and finalizes all results agreed upon by the full committee — no further discussion on those points. Extracts all unresolved points from the previous round (items below threshold), organizes them into a \"dispute list,\" redistributing to all committee members for re-evaluation and re-stance **only on these unresolved items**, ultimately producing this round's decision results.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: Disputes largely converge, producing this round's decision results. **If all committee members approve the decision (no new dispute points), skip the next round and output the final report.** If disputes remain, proceed to the next convergence round.\n\n---\n\n### Final Round: Final Duel\n\n**Special Note**: Regardless of how many rounds are configured, the Final Duel is **always the last round**.\n- 3-round config: Round 3 is the Final Duel\n- 5-round config: Round 5 is the Final Duel\n- N-round config: Round N is the Final Duel\n\n**Action**: After multiple convergence rounds, if deep conflicts remain unresolved, this round is activated. The organizer redistributes these \"deep conflict points\" to all committee members, simplified into binary options — \"Yes/No\" or \"Option A/Option B\" — and requires all committee members to make their final logical stance.\n\n**Pass Condition**: **Consensus Selection** — all judges have identical recommendation order = pass; threshold is NOT a hard requirement.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: Produces the final decision results. Generates the final quantitative voting matrix and automatically produces a 6-section standard decision report.\n\n---\n\n### Round Extension (for non-default round settings)\n\n| Extension Type | Description |\n|:---|:---|\n| Increase rounds (4 or more) | Round 4 follows \"Consensus Sync\"; Round 5 follows \"Final Duel\"; and so on |\n| Decrease rounds (2 rounds) | Round 1 follows \"Independent Evaluation\"; Round 2 directly outputs the final report |\n\n---\n\n## Final Output: 6-Section Consensus Report\n\nThe decision output is a standardized report containing:\n\n1. **Voting Record** — Quantitative scoring matrix\n2. **Summary** — Each committee member's core arguments (with model name noted)\n3. **Execution Plan** — Conclusion version + detailed version\n4. **Unresolved List** — Pending items and responsible parties\n5. **Conclusion Summary** — One-sentence verdict + confidence level\n6. **Risk Advisory** — Key risks and mitigation recommendations\n\n---\n\n## Judgment Method\n\nThree configurable judgment methods (default: unanimous-pass):\n\n| Method | Rule |\n|:---|:---|\n| **Unanimous-pass (default)** | All judges must individually score ≥ threshold |\n| **Average-pass** | Average score of all judges ≥ threshold |\n| **Majority-pass** | More than half of judges score ≥ threshold |\n\n---\n\n## Exception Handling Rules\n\n- **Timeout**: If sub-sessions are not all recovered within 3 minutes, the organizer terminates and returns available results.\n- **Model Failure**: If a judge model fails to respond, the organizer marks it as \"未响应\" and proceeds with available results.\n- **Score Spread**: If max score spread exceeds the convergence threshold, automatically trigger the next round.\n\n---\n*Version: V1.5.2*\n*Developer: Zeekr0808*\n*Email: Zeekr0808@outlook.com*\n\n---\n\n# 多模型决策委员会 — 操作与使用指南\n\n## 版本变更记录\n\n| 版本 | 日期 | 更新内容 |\n|:---|:---|:---|\n| V1.0.0 | 2026-04-19 | 初始版本，包含核心框架、3轮收敛机制、5段式报告 |\n| V1.1.0 | 2026-04-24 | **新增英文操作文档**（与中文合并，英文在上）；经多模型委员会翻译质量审核通过（均分81/100） |\n| V1.1.3 | 2026-04-24 | SKILL.md和references/转为纯中文（精简）；docs/保留双语完整文档（面向全球用户） |\n| V1.2.0 | 2026-04-25 | 恢复为公共ClawHub版本（曾为本地定制版）；同步所有文件 |\n| V1.2.1 | 2026-04-25 | **关键修复**：删除SKILL.md中硬编码的模型列表（A1/A2/A4...）；改为执行时动态扫描通过`openclaw models list`获取实时模型列表；示例中的模型引用改为通用占位符`[模型A/B/C]` |\n| V1.2.2 | 2026-04-25 | 同步：v1.2.1热修复后全文件同步到GitHub和ClawHub |\n| V1.2.3 | 2026-04-25 | **用户体验修复**：明确首次使用流程：自动扫描模型 → 默认前3个 → 提示用户确认或修改评委后再开始决策 |\n| V1.2.4 | 2026-04-25 | **关键修复**：OUTPUT_TEMPLATE.md - 删除所有3轮prompt模板中的「匿名委员」，改为「{模型名称}（实名委员）」格式 |\n| V1.2.5 | 2026-04-25 | **强制判定机制**：SKILL.md - 新增「强制判定规则」章节，明确 Round 1/2 后的阈值判定和收敛判定为不可跳过步骤；禁止行为新增跳项判定禁止规则 |\n| V1.2.6 | 2026-04-25 | **实名显性化**：评委以标准大模型名称显名，透明可溯源；三层架构（组织者/评委/子Agent）；100分制评分；第0轮准备阶段 |\n| V1.5.0 | 2026-04-25 | 同步所有文档版本号至V1.5.0；英文内容全量对齐中文 |\n| V1.5.1 | 2026-04-26 | 新增「每轮开始前通知用户」规则；新增「收敛分差阈值」可配置参数；新增「异常处理规则」章节 |\n| V1.5.2 | 2026-04-26 | 新增「判定方式」可配置参数，明确全票通过制为默认判定规则 |\n| **V1.6.2** | **2026-04-27** | **核心逻辑重构**：取消「收敛」概念，引入「决策点」级别评审；通过判定改为全员通过/非全员通过；「收敛讨论」改为「分歧讨论」；最终报告新增决策点通过状态标注（✅绿色/🟡黄色/🔴红色） |\n| **V1.6.3** | **2026-04-28** | **触发词精确化**：触发条件由长句改为精确词组「多模型决策」「多模型委员会」；**webchat 中间输出适配**：每轮结束后立即向用户展示本轮汇总，不再等待全部轮次结束后统一输出；同步更新所有文档 |\n| **V1.6.5** | **2026-04-30** | **安全加固**：移除 allowed-tools 中未使用的 Exec；最大并发子Agent从13降至6；新增 Write 工具路径限制（仅限 memory 目录）；TROUBLESHOOTING.md 新增模型路由偏差提前终止规则 |\n| **V1.6.6** | **2026-04-30** | **安全修复**：将`openclaw.json`文件扫描改为`openclaw models list`命令，避免暴露 provider token；修正委员会规模描述从3-13为3-6个模型；SKILL.md 新增 Read/Write 工具路径白名单声明 |\n| **V1.6.7** | **2026-04-30** | **可审计性修复**：明确工具使用分工——组织者使用 Read/Write，评委通过 task 参数接收内容且不得调用任何工具；删除\"仅供内部\"标记；消除 OUTPUT_TEMPLATE.md 中评委 spawn 的歧义表述 |\n\n---\n\n## 插件定位\n\n**多模型决策委员会** 是专为 OpenClaw 设计的\"数字智库\"。它支持调动多个不同架构的大模型（如 GPT、Gemini、豆包等）同时对同一任务进行协同思考、对抗辩论与共识合成，有效消除单模型偏见，综合性给出最优建议。\n\n## 核心优势\n\n- **去中心化决策**：交叉验证不同模型逻辑，确保方案的严谨性。\n- **对抗性评审**：通过模型间的互评，快速锁定潜在的逻辑漏洞。\n- **弹性扩容**：根据任务难度，随时增加或减少\"决策委员\"的数量。\n- **消除偏见**：多模型独立评审，避免单一AI的认知盲区。\n- **量化决策**：6维度评分 + 加权矩阵，结论有据可查。\n- **透明可信**：实名委员制、模型身份透明。\n- **灵活配置**：参数可调，如决策成员人数、决策轮次、通过阈值等，均可可自定义。\n\n---\n\n## 核心治理原则：身份纯净\n\n为了确保结果的绝对客观，本插件强制执行 **\"身份纯净原则\"**：\n- **角色透明化**：决策委员会成员均会注明其身份（大模型名称），保持可视化透明。\n- **严禁设定角色**：禁止给委员设定诸如\"架构师\"、\"审计员\"等身份标签。\n- **严禁引导提示**：任务下发时不得针对评审委员设置诱导性提示词或诱导性身份角色，防止产生的偏见。\n- **结论统一原则**：若决策委员会成员根据判断自行调用了子Agent，则必须在本轮结束前汇总子Agent的结果，形成统一结论，再向组织者输出最终结果。\n- **评审不重复原则**：若组织者（即当前使用模型）是评审委员会成员，则组织者提交内容即视同为其在本轮的评审结果，无需重复自评。若组织者不是评审委员会成员，则仅负责组织实施和汇总评委意见，不参与评审。\n\n---\n\n## 使用方法\n-  **启动**：在OpenClaw中，当用户输入的内容包含「多模型决策」或「多模型委员会」时，自动激活多模型决策委员会。首次使用时会提醒用户进入配置模式，并选择模型组合。\n- **配置**：根据需要，可配置参数，如：轮数、阈值、委员数量、委员模型等。\n\n### 参数配置\n配置参数可采用指令方式，如：\"修改配置\"、\"调整参数\"、\"增加委员\"、\"确认配置\"。\n| 指令关键词 | 可调参数 | 说明 |\n|:---|:---|:---|\n| \"换模型\" | 委员模型组合 | 从用户本地接入的模型列表中选择模型组合加入 |\n| \"改轮数\" | 决策轮数 | 默认3轮，日常事务可设为2轮，重大决策可设3-6轮 |\n| \"改阈值\" | 判定阈值 | **通过阈值**：默认≥90%；**待定阈值**：默认[60%, 90%)；**否定阈值**：默认<60% |\n| \"改判定方式\" | 判定方式 | **全票通过** |\n| ~~\"收敛分差阈值\"~~ | ~~收敛分差阈值~~ | ~~已废弃~~ |\n| \"增加委员\" / \"减少委员\" | 委员数量 | 支持2-6个模型同时评审 |\n| \"确认配置\" | 确认当前配置 | 确认当前参数后开始决策 |\n\n## 参数建议\n- **模型选择**：建议至少包含 1 个具备强逻辑推理能力的模型和 1 个具备强中文语境理解能力的模型。\n- **轮次建议**：日常事务建议 2 轮；涉及架构、资金、核心规则的决策，建议 3-5 轮。\n- **阈值设定**：\n  - **严谨型**：≥95%（全票通过）\n  - **效率型**：≥75%（多数通过）\n\n---\n\n## 决策流程（3轮决策机制）\n多模型决策委员会采用 3 轮决策机制，基于决策点级别评审，每轮进行100分制评分，最终给出决策结果。具体流程如下：\n\n---\n\n### 第 0 轮：准备与框架确认 (Preparation & Framework Confirmation)\n**决策准备**：组织者（当前使用模型）接收到决策任务后，将用户背景、需求、待审方案汇总，判断评审方案类型，发起决策申请。内容格式为：\n- **项目名称**：{项目名称}\n- **项目背景**：{项目背景}\n- **项目需求**：{项目需求}\n- **待审方案**：{待审方案}\n- **评审方案类型**：决策点拆分评审\n- **决策点拆分及权重**（仅决策点拆分评审时填写）：\n  - 决策点1（{维度名称}）：{权重}%\n  - 决策点2（{维度名称}）：{权重}%\n  - ……\n- **决策成员**：{决策成员：调取已配置的模型组合}\n- **决策轮次**：{决策轮次：调取已配置的轮数}\n- **阈值设置**：{阈值设置：调取已配置的阈值}\n- **提交决策**：{用户确认评审类型及配置后开始决策}\n\n**评审方案类型**（由组织者根据方案复杂程度判断）：\n- **决策点拆分评审**：复杂/多维度方案，组织者将方案拆解为若干「决策点」，每个决策点对应一个具体评审维度。评委对每个决策点逐项打分（0-100分制）。整体方案得分为各决策点评分的加权平均值，由组织者自动计算，不再由评委单独打分\n\n---\n\n### 第 1 轮：独立评估 (Blind Evaluation)\n\n**动作说明**：组织者接收到决策确认后，将准备流程相关内容分发至各选定模型的独立子会话中。每位委员在**完全无法感知他方意见**的情况下进行背对背评审。\n\n**评审维度**：方案逻辑合理性、完整性、潜在风险分析、实施资源消耗预估、可行性评审、优化建议，并按照100分值进行评分。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：组织者汇总各委员的初评报告及投票结果及打分结果，以及分歧意见或争议点。\n\n---\n\n### 第 2 轮：分歧讨论 (Dispute Discussion)\n\n**动作说明**：组织者将第1轮中标记为「不通过」的决策点整理为「不通过决策点清单」，仅将这些不通过项下发给所有委员，要求评委**只针对这些不通过决策点**进行讨论并重新评分。已通过的决策点不再讨论。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：汇总本轮重新评分结果，若所有决策点全员通过则直接输出最终报告；若仍有不通过项则进入下一轮。\n\n---\n\n### 最后一轮：最终辩论 (Final Duel)\n\n**特殊说明**：无论配置多少轮，最终辩论始终是**最后一轮**。\n- 3轮配置：第3轮为最终辩论\n- 5轮配置：第5轮为最终辩论\n- N轮配置：第N轮为最终辩论\n\n**动作说明**：经过分歧讨论后，仍有决策点不通过时启用。组织者将仍不通过的决策点再次下发给所有委员，要求评委进行最后一轮讨论和重新评分。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：若所有决策点全员通过，输出最终报告；若仍有不通过项，将不通过项标注状态后输出最终报告。\n\n---\n\n### 轮次扩展说明\n\n| 扩展类型 | 说明 |\n|:---|:---|\n| 增加轮次（4轮及以上） | 倒数第2轮参照「分歧讨论」；最后一轮参照「最终辩论」；中间轮次循环分歧讨论 |\n| 减少轮次（2轮） | 第1轮参照\"独立评估\"，第2轮直接输出最终报告 |\n\n---\n\n## 最终输出：5段式共识报告\n\n决策完成后，输出标准化报告，包含：\n\n1. **投票记录** — 量化评分矩阵（含决策点级别评分）\n2. **汇总说明** — 各委员核心论点（注明大模型）\n4. **决策点通过清单** — 各决策点通过状态\n6. **风险提示** — 主要风险与缓解建议\n\n**决策点通过状态标注**：\n\n| 状态 | 含义 | 标注 |\n|:---|:---|:---|\n| 全员通过 | 所有评委在该决策点评分均≥阈值 | ✅ 绿色 — 通过 |\n| 待决策 | 推荐顺序一致但有评委评分低于阈值 | 🟡 黄色 — 待用户确认 |\n| 有分歧 | 推荐顺序不一致 | 🔴 红色 — 有分歧，附原因 |\n\n---\n\n## 判定方式\n\n三种可配置的判定方式（默认：全票通过）：\n\n| 判定方式 | 规则 |\n|:---|:---|\n| **全票通过（默认）** | 所有评委均需≥阈值 |\n\n---\n\n## 异常处理规则\n\n- **超时**：3分钟内未回收所有子会话，组织者终止并返回已收集结果\n- **模型故障**：评委模型无响应时，标记为\"未响应\"，以已收集结果继续进行\n- **超时处理**：3分钟内未回收所有子会话，标记超时评委，以已返回结果继续汇总\n\n---\n*版本: V1.6.7*\n*开发者: Zeekr0808*\n*邮箱: Zeekr0808@outlook.com*\n\nFile v1.8.0:README.md\n\n# Multi-Model Consensus Council — Operation & User Guide\n\n## 📋 Version History\n\n| Version | Date | Changes |\n|:---|:---|:---|\n| V1.0.0 | 2026-04-19 | Initial release: core framework, 3-round convergence, 6-section report |\n| V1.1.0 | 2026-04-24 | Added English operation guide (merged with Chinese, English first); translation reviewed and approved by multi-model committee (Gemini/Doubao/GLM, avg 81/100); wording optimizations |\n| V1.1.3 | 2026-04-24 | SKILL.md and references/ converted to pure Chinese for AI readability; docs/ retains full bilingual documentation for global community |\n| V1.2.0 | 2026-04-25 | Restored public ClawHub version (was local customized); sync all files |\n| V1.2.1 | 2026-04-25 | **CRITICAL FIX**: Removed hardcoded model list (A1/A2/A4...) from SKILL.md; replaced with dynamic scan via `openclaw models list`; examples updated to use generic `[Model A/B/C]` placeholders |\n| V1.2.2 | 2026-04-25 | SYNC: Full file sync to GitHub and ClawHub after v1.2.1 hotfix |\n| V1.2.3 | 2026-04-25 | **UX FIX**: Clarify first-time user flow: auto-scan models → default to first 3 → prompt user to confirm/modify before starting decision |\n| V1.2.4 | 2026-04-25 | **CRITICAL FIX**: OUTPUT_TEMPLATE.md - remove \"匿名委员\" from all 3 round prompt templates; replaced with real-name format `{模型名称}（实名委员）` |\n| V1.2.5 | 2026-04-25 | **ENFORCEMENT**: SKILL.md - add mandatory threshold check rules; add forbidden items for skipping threshold/convergence checks; state machine now has explicit checkpoint enforcement |\n| V1.2.6 | 2026-04-25 | **Real-name Transparency**: Judges use standard model names; 3-layer architecture (organizer/judge/sub-agent); 100-point scoring; Round 0 preparation phase added |\n| V1.5.0 | 2026-04-25 | Sync all document versions to V1.5.0; renamed SCHEMA.md; full English content aligned with Chinese |\n| V1.5.1 | 2026-04-26 | Added round-start user notification rule; added convergence score threshold parameter; added exception handling rules section |\n| V1.5.2 | 2026-04-26 | Added judgment method configurable parameter; clarified unanimous-pass rule as default |\n| **V1.6.0** | **2026-04-27** | **MAJOR FIX**: Rewrote sub-agent result recovery; added runtime environment adaptation matrix; Push-based waiting flow; runtime self-check; simplified state management; added TROUBLESHOOTING.md |\n| **V1.6.1** | **2026-04-27** | **Clarified judgment rules**: Separated \"Pass\" vs \"Convergence\" vs \"Consensus Selection\"; clarified convergence discussion scope; renamed \"Round 3\" to \"Final Round\"; added Consensus Selection for final round |\n| **V1.6.2** | **2026-04-27** | **Core logic refactored**: Removed \"Convergence\"; introduced \"Decision Point\" level review; \"Convergence Discussion\" renamed to \"Dispute Discussion\"; final report includes decision point status (✅/🟡/🔴) |\n| **V1.6.5** | **2026-04-30** | **Security Hardening**: Removed unused `Exec` from allowed-tools; reduced max concurrent sub-agents from 13 to 6; added `Write` tool path restriction (memory directory only); added model-routing early termination rule in TROUBLESHOOTING.md |\n| **V1.6.6** | **2026-04-30** | **Security Fix**: Replaced `openclaw.json` file scan with `openclaw models list` command to avoid exposing provider tokens; corrected committee size description from 3-13 to 3-6 models; added Read/Write tool path whitelist in SKILL.md |\n| **V1.6.7** | **2026-04-30** | **Auditability Fix**: Clarified tool usage scope—organizers use Read/Write, judges receive content via task parameter and must not call any tools; removed contradictory \"internal-only\" markings; eliminated ambiguous judge spawn instructions in OUTPUT_TEMPLATE.md |\n\n---\n\n## 🎯 Plugin Overview\n\n**Multi-Model Consensus Council** is a \"Digital Think Tank\" designed for OpenClaw. It coordinates multiple large language models (e.g., GPT, Gemini, Doubao) to collaboratively analyze, debate, and synthesize decisions, effectively eliminating single-model bias.\n\n### Core Advantages\n\n- **Decentralized Decision-Making**: Cross-validates reasoning across models, ensuring rigor.\n- **Adversarial Review**: Rapidly identifies logical flaws through peer evaluation.\n- **Elastic Scaling**: Adjust the number of \"decision committee members\" anytime based on task complexity.\n- **Bias Elimination**: Multi-model independent review avoids single-AI cognitive blind spots.\n- **Quantitative Decision-Making**: 6-dimension scoring + weighted matrix, conclusions are evidence-based.\n- **Transparency & Trust**: Real-name committee system, model identity visible.\n- **Flexible Configuration**: All parameters adjustable — number of members, decision rounds, pass threshold, etc.\n\n---\n\n## ⚖️ Core Governance Principle: Identity Purity\n\nTo ensure absolute objectivity, this plugin enforces the **\"Identity Purity Principle\"**:\n\n- **Identity Transparency**: All committee members are identified by their model name (e.g., Kimi K2.6, GPT-5.4), maintaining visual transparency.\n- **No Role Assignment**: Committee members are forbidden from having preset roles such as \"architect\" or \"auditor\".\n- **No Biased Prompting**: Tasks must not include inducing prompts or identity roles targeted at committee members, preventing bias from prompt engineering.\n- **No Sub-Agent Delegation**: Judges are strictly prohibited from spawning sub-agents. They must think independently and output their own conclusions. Complex tasks are decomposed and distributed by the organizer.\n- **No Duplicate Review Principle**: If the organizer (the current model in use) is a committee member, their submitted content counts as their Round 1 review result — no duplicate self-review is required. If the organizer is not a committee member, they are responsible only for organizing and summarizing; they do not participate in the review.\n\n---\n\n## How to Use\n\n- **Activation**: In OpenClaw, type \"启动多模型决策委员会\" (Start Multi-Model Consensus Council), \"审议这个方案\" (Review this plan), \"对这个方案进行决策\" (Make a decision on this plan), or \"启动决策\" (Start decision) to activate. On first use, the system will prompt you to enter configuration mode and select the model lineup.\n- **Configuration**: Adjust parameters as needed, such as rounds, threshold, number of members, and member models.\n\n### Parameter Configuration\n\n| Command Keyword | Adjustable Parameter | Description |\n|:---|:---|:---|\n| \"换模型\" / \"Change models\" | Committee model mix | Select models from the locally connected model list |\n| \"改轮数\" / \"Change rounds\" | Decision rounds | Default 3 rounds; 2 for routine tasks; 3-6 for critical decisions |\n| \"改阈值\" / \"Change threshold\" | Judgment thresholds | **Pass threshold**: default ≥90%; **Pending threshold**: default [72%, 90%); **Reject threshold**: default <72% |\n| \"增加委员\" / \"Add member\" | Committee size | Increase number of committee members |\n| \"减少委员\" / \"Remove member\" | Committee size | Decrease number of committee members |\n| \"确认配置\" / \"Confirm config\" | Confirm current config | Confirm current parameters and begin decision |\n\n### Parameter Recommendations\n\n- **Model Selection**: Recommend at least 1 model with strong logical reasoning and 1 model with strong Chinese-language comprehension.\n- **Round Recommendations**: Routine tasks: 2 rounds; Decisions involving architecture, funding, or core rules: 3-5 rounds.\n- **Threshold Settings**:\n  - **Strict**: ≥95% (near-unanimous)\n  - **Efficient**: ≥75% (majority)\n\n---\n\n## Decision Process (3-Round Convergence Mechanism)\n\nThe Multi-Model Consensus Council adopts a 3-round convergence mechanism, with each round using a **100-point scoring system**, ultimately producing a decision result.\n\n---\n\n### Round 0: Preparation (Preparation)\n\n**Task**: Upon receiving a decision task, the **organizer** (the current model in use) summarizes the user background, requirements, and the proposal to initiate the decision application. Content format:\n\n- **Project Name**: {Project Name}\n- **Project Background**: {Project Background}\n- **Project Requirements**: {Project Requirements}\n- **Proposal Under Review**: {Proposal Under Review}\n- **Decision Members**: {Decision Members: pull from configured model lineup}\n- **Decision Rounds**: {Decision Rounds: pull from configured round count}\n- **Threshold Settings**: {Threshold Settings: pull from configured threshold}\n- **Submit Decision**: {Submit whether the user wants to begin decision}\n\n---\n\n### Round 1: Independent Evaluation (Blind Evaluation)\n\n**Action**: Upon receiving the decision confirmation, the organizer distributes the preparation content to each selected model's independent sub-session. Each committee member conducts closed-door review **without any awareness of other members' opinions**.\n\n**Review Dimensions**: Logical completeness, risk analysis, resource cost estimation, feasibility review, optimization suggestions — all scored on a **100-point scale**.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: The organizer summarizes each committee member's initial review report, voting results, scores, and disagreement points or controversy areas.\n\n---\n\n### Round 2: Dispute Discussion\n\n**Action**: The organizer extracts all decision points that did NOT achieve full pass in Round 1 (any judge scoring below threshold) and compiles them into an \"unresolved decision points list.\" These unresolved points only are redistributed to all committee members for re-discussion and re-scoring. Points already with full pass are frozen and not re-discussed.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: Summarizes re-scoring results. If all previously unresolved points now achieve full pass, output the final report. If any remain unresolved, proceed to the final round.\n\n---\n\n### Final Round: Final Duel\n\n**Special Note**: Regardless of how many rounds are configured, the Final Duel is **always the last round**.\n- 3-round config: Round 3 is the Final Duel\n- 5-round config: Round 5 is the Final Duel\n- N-round config: Round N is the Final Duel\n\n**Action**: After dispute discussion, if any decision points still have judges scoring below threshold, those unresolved points are redistributed to all committee members for a final round of discussion and re-scoring.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: If all decision points now achieve full pass, output the final report. If any remain unresolved, output the final report with each decision point's status tagged: ✅ Full Pass / 🟡 Pending User Decision / 🔴 Disputed.\n\n---\n\n### Round Extension (for non-default round settings)\n\n| Extension Type | Description |\n|:---|:---|\n| Increase rounds (4 or more) | 2nd-to-last round follows \"Dispute Discussion\"; last round follows \"Final Duel\"; middle rounds loop dispute discussion |\n| Decrease rounds (2 rounds) | Round 1 follows \"Independent Evaluation\"; Round 2 directly outputs the final report (treated as final round) |\n\n---\n\n## Final Output: 6-Section Consensus Report\n\nThe decision output is a standardized report containing:\n\n1. **Voting Record** — Quantitative scoring matrix (by decision point)\n2. **Summary** — Each committee member's core arguments (with model name noted)\n3. **Execution Plan** — Conclusion version + detailed version\n4. **Decision Point Status** — Each decision point's pass status\n5. **Conclusion Summary** — One-sentence verdict + confidence level\n6. **Risk Advisory** — Key risks and mitigation recommendations\n\n**Decision Point Status Tags**:\n| Status | Meaning | Tag |\n|:---|:---|:---|\n| Full Pass | All judges scored ≥ threshold on this point | ✅ Green — Passed |\n| Pending Decision | All judges chose same option but some scores below threshold | 🟡 Yellow — Pending User Confirmation |\n| Disputed | Judges chose different options | 🔴 Red — Disputed, reason noted |\n\n---\n*Version: V1.6.7*\n*Developer: Zeekr0808*\n*Email: Zeekr0808@outlook.com*\n\n---\n\n# 多模型决策委员会 — 操作与使用指南\n\n## 版本变更记录\n\n| 版本 | 日期 | 更新内容 |\n|:---|:---|:---|\n| V1.0.0 | 2026-04-19 | 初始版本，包含核心框架、3轮收敛机制、5段式报告 |\n| V1.1.0 | 2026-04-24 | **新增英文操作文档**（与中文合并，英文在上）；经多模型委员会翻译质量审核通过（均分81/100） |\n| V1.1.3 | 2026-04-24 | SKILL.md和references/转为纯中文（精简）；docs/保留双语完整文档（面向全球用户） |\n| V1.2.0 | 2026-04-25 | 恢复为公共ClawHub版本（曾为本地定制版）；同步所有文件 |\n| V1.2.1 | 2026-04-25 | **关键修复**：删除SKILL.md中硬编码的模型列表（A1/A2/A4...）；改为执行时动态扫描通过`openclaw models list`获取实时模型列表；示例中的模型引用改为通用占位符`[模型A/B/C]` |\n| V1.2.2 | 2026-04-25 | 同步：v1.2.1热修复后全文件同步到GitHub和ClawHub |\n| V1.2.3 | 2026-04-25 | **用户体验修复**：明确首次使用流程：自动扫描模型 → 默认前3个 → 提示用户确认或修改评委后再开始决策 |\n| V1.2.4 | 2026-04-25 | **关键修复**：OUTPUT_TEMPLATE.md - 删除所有3轮prompt模板中的「匿名委员」，改为「{模型名称}（实名委员）」格式 |\n| V1.2.5 | 2026-04-25 | **强制判定机制**：SKILL.md - 新增「强制判定规则」章节，明确 Round 1/2 后的阈值判定和收敛判定为不可跳过步骤；禁止行为新增跳项判定禁止规则 |\n| V1.2.6 | 2026-04-25 | **实名显性化**：评委以标准大模型名称显名，透明可溯源；三层架构（组织者/评委/子Agent）；100分制评分；第0轮准备阶段 |\n| V1.5.0 | 2026-04-25 | 同步所有文档版本号至V1.5.0；英文内容全量对齐中文 |\n| V1.5.1 | 2026-04-26 | 新增「每轮开始前通知用户」规则；新增「收敛分差阈值」可配置参数；新增「异常处理规则」章节 |\n| V1.5.2 | 2026-04-26 | 新增「判定方式」可配置参数，明确全票通过制为默认判定规则 |\n| **V1.6.0** | **2026-04-27** | **重大修复**：重写子Agent结果回收机制，新增运行时环境适配矩阵；新增Push-based等待流程；新增运行时自检；简化状态管理；新增TROUBLESHOOTING.md |\n| **V1.6.1** | **2026-04-27** | **判定规则澄清**：区分「通过」「收敛」「一致性选择」三个独立概念；澄清收敛讨论范围；「第3轮」更名为「最后一轮」；新增最终轮「一致性选择」机制 |\n| **V1.6.2** | **2026-04-27** | **核心逻辑重构**：取消「收敛」概念，引入「决策点」级别评审；通过判定改为全员通过/非全员通过；「收敛讨论」改为「分歧讨论」；最终报告新增决策点通过状态标注（✅绿色/🟡黄色/🔴红色） |\n| **V1.6.5** | **2026-04-30** | **安全加固**：移除 allowed-tools 中未使用的 Exec；最大并发子Agent从13降至6；新增 Write 工具路径限制（仅限 memory 目录）；TROUBLESHOOTING.md 新增模型路由偏差提前终止规则 |\n| **V1.6.6** | **2026-04-30** | **安全修复**：将`openclaw.json`文件扫描改为`openclaw models list`命令，避免暴露 provider token；修正委员会规模描述从3-13为3-6个模型；SKILL.md 新增 Read/Write 工具路径白名单声明 |\n| **V1.6.7** | **2026-04-30** | **可审计性修复**：明确工具使用分工——组织者使用 Read/Write，评委通过 task 参数接收内容且不得调用任何工具；删除\"仅供内部\"标记；消除 OUTPUT_TEMPLATE.md 中评委 spawn 的歧义表述 |\n\n---\n\n## 插件定位\n\n**多模型决策委员会** 是专为 OpenClaw 设计的\"数字智库\"。它支持调动多个不同架构的大模型（如 GPT、Gemini、豆包等）同时对同一任务进行协同思考、对抗辩论与共识合成，有效消除单模型偏见，综合性给出最优建议。\n\n## 核心优势\n\n- **去中心化决策**：交叉验证不同模型逻辑，确保方案的严谨性。\n- **对抗性评审**：通过模型间的互评，快速锁定潜在的逻辑漏洞。\n- **弹性扩容**：根据任务难度，随时增加或减少\"决策委员\"的数量。\n- **消除偏见**：多模型独立评审，避免单一AI的认知盲区。\n- **量化决策**：6维度评分 + 加权矩阵，结论有据可查。\n- **透明可信**：实名委员制、模型身份透明。\n- **灵活配置**：参数可调，如决策成员人数、决策轮次、通过阈值等，均可可自定义。\n\n---\n\n## 核心治理原则：身份纯净\n\n为了确保结果的绝对客观，本插件强制执行 **\"身份纯净原则\"**：\n- **角色透明化**：决策委员会成员均会注明其身份（大模型名称），保持可视化透明。\n- **严禁设定角色**：禁止给委员设定诸如\"架构师\"、\"审计员\"等身份标签。\n- **严禁引导提示**：任务下发时不得针对评审委员设置诱导性提示词或诱导性身份角色，防止产生的偏见。\n- **严禁子Agent委托**：评委严禁spawn任何子Agent，只能独立思考和输出结论；禁止二次委托，所有子任务由组织者统一分发和回收，确保链路完全可控。\n- **评审不重复原则**：若组织者（即当前使用模型）是评审委员会成员，则组织者提交内容即视同为其在本轮的评审结果，无需重复自评。若组织者不是评审委员会成员，则仅负责组织实施和汇总评委意见，不参与评审。\n\n---\n\n## 使用方法\n-  **启动**：在OpenClaw中，输入指令\"启动多模型决策委员会\"、\"审议这个方案\"、\"对这个方案进行决策\"、\"启动决策\"即可启动多模型决策委员会。首次使用时会提醒用户进入配置模式，并选择模型组合。\n- **配置**：根据需要，可配置参数，如：轮数、阈值、委员数量、委员模型等。\n\n### 参数配置\n配置参数可采用指令方式，如：\"修改配置\"、\"调整参数\"、\"增加委员\"、\"确认配置\"。\n| 指令关键词 | 可调参数 | 说明 |\n|:---|:---|:---|\n| \"换模型\" | 委员模型组合 | 从用户本地接入的模型列表中选择模型组合加入 |\n| \"改轮数\" | 决策轮数 | 默认3轮，日常事务可设为2轮，重大决策可设3-6轮 |\n| \"改阈值\" | 判定阈值 | **通过阈值**：默认≥90%；**待定阈值**：默认[60%, 90%)；**否定阈值**：默认<60% |\n| \"增加委员\" / \"减少委员\" | 委员数量 | 支持2-6个模型同时评审 |\n| \"确认配置\" | 确认当前配置 | 确认当前参数后开始决策 |\n\n## 参数建议\n- **模型选择**：建议至少包含 1 个具备强逻辑推理能力的模型和 1 个具备强中文语境理解能力的模型。\n- **轮次建议**：日常事务建议 2 轮；涉及架构、资金、核心规则的决策，建议 3-5 轮。\n- **阈值设定**：\n  - **严谨型**：≥95%（全票通过）\n  - **效率型**：≥75%（多数通过）\n\n---\n\n## 决策流程（3轮决策机制）\n多模型决策委员会采用 3 轮决策机制，基于决策点级别评审，每轮进行100分制评分，最终给出决策结果。具体流程如下：\n\n---\n\n### 第 0 轮：准备与框架确认 (Preparation & Framework Confirmation)\n**决策准备**：组织者（当前使用模型）接收到决策任务后，将用户背景、需求、待审方案汇总，判断评审方案类型，发起决策申请。内容格式为：\n- **项目名称**：{项目名称}\n- **项目背景**：{项目背景}\n- **项目需求**：{项目需求}\n- **待审方案**：{待审方案}\n- **评审方案类型**：决策点拆分评审\n- **决策点拆分及权重**（仅决策点拆分评审时填写）：\n  - 决策点1（{维度名称}）：{权重}%\n  - 决策点2（{维度名称}）：{权重}%\n  - ……\n- **决策成员**：{决策成员：调取已配置的模型组合}\n- **决策轮次**：{决策轮次：调取已配置的轮数}\n- **阈值设置**：{阈值设置：调取已配置的阈值}\n- **提交决策**：{用户确认评审类型及配置后开始决策}\n\n**评审方案类型**（由组织者根据方案复杂程度判断）：\n- **决策点拆分评审**：复杂/多维度方案，组织者将方案拆解为若干「决策点」，每个决策点对应一个具体评审维度。评委对每个决策点逐项打分（0-100分制）。整体方案得分为各决策点评分的加权平均值，由组织者自动计算，不再由评委单独打分\n\n---\n\n### 第 1 轮：独立评估 (Blind Evaluation)\n\n**动作说明**：组织者接收到决策确认后，将准备流程相关内容分发至各选定模型的独立子会话中。每位委员在**完全无法感知他方意见**的情况下进行背对背评审。\n\n**评审维度**：方案逻辑合理性、完整性、潜在风险分析、实施资源消耗预估、可行性评审、优化建议，并按照100分值进行评分。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：组织者汇总各委员的初评报告及投票结果及打分结果，以及分歧意见或争议点。\n\n---\n\n### 第 2 轮：分歧讨论 (Dispute Discussion)\n\n**动作说明**：组织者将第1轮中标记为「不通过」的决策点整理为「不通过决策点清单」，仅将这些不通过项下发给所有委员，要求评委**只针对这些不通过决策点**进行讨论并重新评分。已通过的决策点不再讨论。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：汇总本轮重新评分结果，若所有决策点全员通过则直接输出最终报告；若仍有不通过项则进入下一轮。\n\n---\n\n### 最后一轮：最终辩论 (Final Duel)\n\n**特殊说明**：无论配置多少轮，最终辩论始终是**最后一轮**。\n- 3轮配置：第3轮为最终辩论\n- 5轮配置：第5轮为最终辩论\n- N轮配置：第N轮为最终辩论\n\n**动作说明**：经过分歧讨论后，仍有决策点不通过时启用。组织者将仍不通过的决策点再次下发给所有委员，要求评委进行最后一轮讨论和重新评分。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：若所有决策点全员通过，输出最终报告；若仍有不通过项，将不通过项标注状态后输出最终报告。\n\n---\n\n### 轮次扩展说明\n\n| 扩展类型 | 说明 |\n|:---|:---|\n| 增加轮次（4轮及以上） | 倒数第2轮参照「分歧讨论」；最后一轮参照「最终辩论」；中间轮次循环分歧讨论 |\n| 减少轮次（2轮） | 第1轮参照\"独立评估\"，第2轮直接输出最终报告 |\n\n---\n\n## 最终输出：5段式共识报告\n\n决策完成后，输出标准化报告，包含：\n\n1. **投票记录** — 量化评分矩阵（含决策点级别评分）\n2. **汇总说明** — 各委员核心论点（注明大模型）\n4. **决策点通过清单** — 各决策点通过状态\n6. **风险提示** — 主要风险与缓解建议\n\n**决策点通过状态标注**：\n\n| 状态 | 含义 | 标注 |\n|:---|:---|:---|\n| 全员通过 | 所有评委在该决策点评分均≥阈值 | ✅ 绿色 — 通过 |\n| 待决策 | 推荐顺序一致但有评委评分低于阈值 | 🟡 黄色 — 待用户确认 |\n| 有分歧 | 推荐顺序不一致 | 🔴 红色 — 有分歧，附原因 |\n\n---\n*版本: V1.6.7*\n*开发者: Zeekr0808*\n*邮箱: Zeekr0808@outlook.com*\n\nFile v1.8.0:_meta.json\n\n{\n  \"ownerId\": \"kn7e3gt0d3qkc3nmkhqzcx15ks85fz9s\",\n  \"slug\": \"multi-model-consensus\",\n  \"version\": \"1.8.0\",\n  \"publishedAt\": 1778942873996\n}\n\nFile v1.8.0:references/OUTPUT_TEMPLATE.md\n\n# 多模型决策委员会 — 输出模板\n\n> 本文件定义各阶段标准模板。T2/T4/T6 为评委分发 Prompt，已在末尾附录；T1/T3/T5/T7 为本文件正文模板。\n\n---\n\n## T1 — 决策准备确认模板（第0轮）\n\n组织者向用户确认决策框架，**须完整填写以下所有字段**：\n\n```\n═══════════════════════════════════════════════\n🏛️ 多模型决策委员会 — 决策准备确认\n═══════════════════════════════════════════════\n\n【项目名称】\n（从用户输入提炼，不超过20字）\n\n【项目背景】\n（描述当前情境、痛点或决策缘由，50字以内）\n\n【项目需求】\n（用户期望解决的核心问题，50字以内）\n\n【待审方案】\n（完整呈现待评审方案内容）\n\n───────────────────────────────────────────────\n【评审规则】：各位评委背对背独立评审，由组织者汇总。若所有评委评审项均达到通过阈值，该项视为**通过**；若所有评委评审项均未达到否定阈值，该项视为**不通过**；其余情况视为**待定**。\n\n【决策点拆分及权重】\n  决策点1（{维度名称}）：{权重}%\n  决策点2（{维度名称}）：{权重}%\n  ……\n  （权重总和须等于100%）\n\n───────────────────────────────────────────────\n【评委配置】\n| 序号 | 委员 | 模型 | 备注 |\n|:---:|:---|:---|:---|\n| 1 | 评委1 | {模型名}（{Provider}） | |\n| 2 | 评委2 | {模型名}（{Provider}） | |\n| 3 | 评委3 | {模型名}（{Provider}） | |\n\n【决策轮次】：{3} 轮\n【阈值设置】：通过≥90% / 待定[60%,90%) / 否定<60%\n【判定方式】：全票通过\n\n是否需要修改「评委」、「决策轮次」、「通过阈值」、「判定方式」等配置？若有需要，请告诉我。\n\n═══════════════════════════════════════════════\n请确认以上框架后，回复「确认」或「开始决策」。\n═══════════════════════════════════════════════\n```\n\n---\n\n## T3 — 第1轮汇总模板\n\n第1轮独立评估结束后，**立即输出以下内容**：\n\n```\n═══════════════════════════════════════════════\n📊 第 1 轮独立评估 — 评分汇总\n═══════════════════════════════════════════════\n\n【评审类型】：决策点拆分评审\n【评委到齐】：{X}/{X}（{缺席评委名} 超时未提交）\n\n───────────────────────────────────────────────\n【本轮判定结果】：{通过 / 待定 / 不通过}\n阈值设置：通过≥90% / 待定[60%,90%) / 否定<60%\n═══════════════════════════════════════════════\n```\n\n### 第一轮决策项目展示\n\n**决策点拆分评审**（评委对每个决策点评分，加权平均）：\n\n| 决策点（权重） | {决策点1}（{X}%） | {决策点2}（{X}%） | … | 加权后分数 |\n|:---|:---:|:---:|:---:|:---:|\n| {模型A}（实名） | {X}/100 | {X}/100 | … | **{X}/100** |\n| {模型B}（实名） | {X}/100 | {X}/100 | … | **{X}/100** |\n| **评委均分** | {X}/100 | {X}/100 | … | **{X}/100** |\n| **通过情况** | {✅/🟡/🔴} | {✅/🟡/🔴} | … | — |\n\n───────────────────────────────────────────────\n已通过决策点已封存，不再进入下轮讨论。\n本轮🟡待定项：{决策点1}（{X}%）……由组织者综合评委建议，形成调整优化方向后，进入第二轮分歧讨论。\n本轮🔴不通过项：{决策点2}（{X}%）……由组织者综合评委建议，形成调整优化方向后，进入第二轮分歧讨论。\n请问是否「开始」或「继续」进入第二轮讨论？还是「结束」，直接结束讨论，直接输出最终结果。您可以说「开始」或「继续」，也可以说「结束」。\n\n---\n\n## T5 — 第2轮汇总模板\n\n第2轮分歧讨论结束后，**立即输出以下内容**：\n\n```\n═══════════════════════════════════════════════\n📊 第 2 轮分歧讨论 — 评分汇总\n═══════════════════════════════════════════════\n\n【本轮讨论范围】：仅第 1 轮🟡待定及🔴不通过决策点\n【评委到齐】：{X}/{X}（{缺席评委名} 超时未提交）\n\n───────────────────────────────────────────────\n【本轮判定结果】：{通过 / 待定 / 不通过}\n阈值设置：通过≥90% / 待定[60%,90%) / 否定<60%\n═══════════════════════════════════════════════\n```\n\n### 第二轮决策项目展示\n\n**不通过决策点本轮评分**：\n\n| 决策点 | 评委 | 第1轮得分 | 本轮得分 | 变化 | 是否达标 |\n|:---|:---|:---:|:---:|:---:|:---:|\n| {决策点1} | {模型A}（实名） | {X} | {X} | {+/-X} | {✅/❌} |\n| {决策点1} | {模型B}（实名） | {X} | {X} | {+/-X} | {✅/❌} |\n| {决策点2} | {模型A}（实名） | {X} | {X} | {+/-X} | {✅/❌} |\n| {决策点2} | {模型B}（实名） | {X} | {X} | {+/-X} | {✅/❌} |\n\n**评委立场变化追踪**：\n\n| 评委 | 决策点1立场变化 | 决策点2立场变化 |\n|:---|:---|:---|\n| {模型A}（实名） | {坚持/改变/部分同意} | {坚持/改变/部分同意} |\n| {模型B}（实名） | {坚持/改变/部分同意} | {坚持/改变/部分同意} |\n\n───────────────────────────────────────────────\n已通过决策点已封存，不再进入下轮讨论。\n本轮🟡待定项：{决策点1}（{X}%）……由组织者综合评委建议，形成调整优化方向后，进入第三轮（如有）或最终报告。\n本轮🔴不通过项：{决策点2}（{X}%）……由组织者综合评委建议，形成调整优化方向后，进入第三轮（如有）或最终报告。\n请问是否「继续」进入下一轮讨论？还是「结束」，直接输出最终结果。您可以说「继续」或「结束」。\n\n═══════════════════════════════════════════════\n```\n\n---\n\n## T7 — 最终报告模板（5段式共识报告）\n\n决策全部轮次完成后，**输出标准化报告**：\n\n```\n═══════════════════════════════════════════════\n🏛️ 多模型决策共识报告\n═══════════════════════════════════════════════\n\n【决策主题】：（用户输入的决策任务简述，不超过20字）\n【生成时间】：（自动生成时间戳，格式：YYYY-MM-DD HH:MM）\n【评委配置】：{模型A}（{Provider}）、{模型B}（{Provider}）、{模型C}（{Provider}）\n【实际执行轮数】：{X} 轮（{轮次明细}）\n【阈值设置】：通过≥90% / 待定[60%,90%) / 否定<60%\n【判定方式】：全票通过\n\n═══════════════════════════════════════════════\n一、投票记录\n═══════════════════════════════════════════════\n\n【评审类型】：决策点拆分评审\n\n（对应矩阵格式，参见T3模板说明）\n\n═══════════════════════════════════════════════\n二、汇总说明\n═══════════════════════════════════════════════\n\n### {模型A}（实名） — 核心论点\n\n（详细说明该评委的分析逻辑、评分依据、主要担忧，100字以内）\n\n### {模型B}（实名） — 核心论点\n\n（详细说明该评委的分析逻辑、评分依据、主要担忧，100字以内）\n\n### {模型C}（实名） — 核心论点\n\n（详细说明该评委的分析逻辑、评分依据、主要担忧，100字以内）\n\n\n三、决策点通过清单\n═══════════════════════════════════════════════\n\n| 决策点 | 权重 | 最终评分 | 状态 | 说明 |\n|:---|:---:|:---:|:---:|:---|\n| {决策点1} | {X}% | {X}/100 | ✅ 通过 | {无/有保留意见} |\n| {决策点2} | {X}% | {X}/100 | 🟡 待决策 | {评委{X}评分低于阈值，建议用户确认} |\n| {决策点3} | {X}% | {X}/100 | 🔴 有分歧 | {评委{X}与评委{Y}结论相反} |\n\n**通过状态说明**：\n- ✅ 绿色 — 所有评委该决策点评分均≥阈值，全员通过\n- 🟡 黄色 — 推荐顺序一致但有评委评分低于阈值，建议用户确认\n- 🔴 红色 — 推荐顺序不一致，存在实质分歧\n\n═══════════════════════════════════════════════\n四、结论总结\n═══════════════════════════════════════════════\n\n**总体结论**：（综合各轮评审结果，给出最终决策结论，100字以内）\n\n**置信度**：{X}%（基于{X}轮评审，{X}/{X}评委推荐）\n\n**主要分歧**：（如有不通过决策点，详细描述分歧焦点，50字以内）\n\n**建议内容**：（具体可执行的建议，50字以内）\n\n═══════════════════════════════════════════════\n五、风险提示\n═══════════════════════════════════════════════\n\n**主要风险**：（最重要的1个风险，20字以内）\n\n| 风险类型 | 风险描述 | 概率 | 影响 | 缓解措施 |\n|:---|:---|:---:|:---:|:---|\n| 技术风险 | {描述，不超过20字} | {高/中/低} | {高/中/低} | {措施，不超过15字} |\n| 进度风险 | {描述，不超过20字} | {高/中/低} | {高/中/低} | {措施，不超过15字} |\n| 资源风险 | {描述，不超过20字} | {高/中/低} | {高/中/低} | {措施，不超过15字} |\n\n═══════════════════════════════════════════════\n附录：决策过程回顾\n═══════════════════════════════════════════════\n\n- **第 1 轮完成时间**：（时间戳）\n- **第 2 轮完成时间**：（时间戳）\n- **最终辩论完成时间**：（时间戳，如有）\n- **不通过决策点（第1轮）**：{决策点1、决策点2}\n- **不通过决策点（第2轮）**：{决策点1}\n- **最终不通过决策点**：{决策点1}（标注🔴/🟡状态）\n- **超时评委**：{评委名}（如有）\n- **异常记录**：{如有}\n\n═══════════════════════════════════════════════\n🏛️ 报告生成完毕 | 置信度：{X}%\n═══════════════════════════════════════════════\n```\n\n---\n\n## 附录：评委分发 Prompt 模板（T2 / T4 / T6）\n\n> 以下为组织者分发给各评委的评审任务 Prompt，**由组织者根据轮次填充后，通过 sessions_spawn 发送给各评委子会话**。\n\n### T2 — 第1轮独立评审 Prompt\n\n```\n【角色】你是多模型决策委员会的评委委员，你的身份是：{标准大模型名称}（实名委员）。\n【任务】独立评审以下方案，不与任何其他评委交流。\n       【硬性约束】评委严禁spawn子Agent或调用任何外部工具，必须基于自身知识独立完成评审。\n【评审方案】\n（用户提交的方案内容）\n\n【评审维度】（6维度，每维度100分制）\n1. 逻辑完整性：（方案推理是否严密）— X/100\n2. 风险控制：（潜在风险识别与应对）— X/100\n3. 实施资源：（成本、人力、时间估算）— X/100\n4. 可行性：（技术与业务可行性）— X/100\n5. 长期价值：（战略意义与可持续性）— X/100\n6. 用户接受度：（用户体验与采用意愿）— X/100\n【硬约束】你只能输出一套评分。\n         禁止在同一输出中出现两套或以上相互独立的评分体系。\n         若涉及多个视角，请在「主要担忧」中简要注明，不要拆分为多个独立评分。\n\n【输出格式】\n决策点评分：[每个决策点0-100分，列出所有决策点]\n推荐顺序：[方案A优先 / 方案B优先 / 持平]\n核心论点：[200字以内的主要理由]\n主要担忧：[100字以内的主要风险点]\n```\n\n### T4 — 第2轮分歧讨论 Prompt\n\n```\n【角色】你是多模型决策委员会的评委委员（实名：{标准大模型名称}）。\n【背景】第1轮独立评审已完成，以下是各评委对各决策点的评分结果（实名）：\n（汇总第1轮决策点评分矩阵，注明各评委真实模型名称）\n\n【不通过决策点清单】\n（列出第1轮中未达到全员通过的决策点，即：任一评委评分<阈值的决策点）\n\n【任务】请阅读其他评委的观点后，**仅针对以上不通过决策点**进行二次表态和重新评分。\n- 若同意某评委的观点，请说明理由\n- 若坚持自己的观点，请提供更强论据\n- 若改变立场，明确说明原因\n\n【输出格式】\n对分歧点1的态度：[坚持/改变/部分同意] — 理由（50字以内）\n对分歧点2的态度：[坚持/改变/部分同意] — 理由（50字以内）\n决策点重新评分：[每个不通过决策点重新打分，0-100分]\n最终推荐顺序：[是否与第1轮一致]\n```\n\n### T6 — 最后一轮最终辩论 Prompt\n\n```\n【角色】你是多模型决策委员会的评委委员（实名：{标准大模型名称}）。\n【背景】经过第2轮分歧讨论后，仍存在以下不通过的决策点（实名评委）：\n（列出仍未解决的分歧点，标注各评委真实身份）\n\n【任务】请阅读其他评委的观点后，**仅针对以上不通过决策点**进行最终表态和重新评分。\n- 若同意某评委的观点，请说明理由\n- 若坚持自己的观点，请提供更强论据\n- 若改变立场，明确说明原因\n\n【输出格式】\n对分歧点1的态度：[坚持/改变/部分同意] — 理由（50字以内）\n对分歧点2的态度：[坚持/改变/部分同意] — 理由（50字以内）\n决策点重新评分：[每个不通过决策点重新打分，0-100分]\n最终推荐顺序：[是否与第2轮一致]\n```\n\nFile v1.8.0:references/SCHEMA.md\n\n# 多模型决策委员会 — 状态文件 Schema（已归档）\n\n> ⚠️ **V1.6.0 说明**：本文档描述的状态文件 Schema 已归档。V1.6.0 起不再使用持久化状态文件，改用「轻量级状态跟踪」（会话上下文变量）。本文件保留供历史参考。\n\n## 状态文件字段\n\n| 字段 | 类型 | 必填 | 说明 |\n|:---|:---|:---:|:---|\n| `state` | string | ✅ | 当前状态：IDLE / ROUND_0 / ROUND_1 / ROUND_2 / ROUND_3 / COMPLETE |\n| `config` | object | ✅ | 评委配置 |\n| `config.models` | array | ✅ | 评委模型列表，如 [\"模型A\", \"模型B\", \"模型C\"] |\n| `config.rounds` | integer | ✅ | 决策轮数，2-6，默认3 |\n| `config.threshold` | number | ✅ | 通过阈值，0-1，如 0.9 |\n| `round0` | object | — | 准备阶段汇总内容 |\n| `round1` | object | — | 第1轮结果（含 judges[].used_subagent） |\n| `round2` | object | — | 第2轮结果 |\n| `round3` | object | — | 第3轮结果（仅在有分歧时） |\n| `report` | object | — | 最终5段式共识报告 |\n\n---\n\n## 版本演进规则\n\n1. **不得单方面升版**：任何 Schema 变更必须经过委员会评审通过\n2. **向后兼容**：新版本应尽可能兼容旧版本数据格式\n3. **升版记录**：每次升版须在 CHANGELOG.md 中详细记录变更内容\n\nFile v1.8.0:references/STATE_MACHINE.md\n\n# 多模型决策委员会 — 状态机规范（概念参考）\n\n> ⚠️ **V1.6.0 说明**：本文档描述的状态机逻辑仍为概念规范，但实际执行时已**不再使用持久化状态文件**。组织者改用「轻量级状态跟踪」（会话上下文变量）来跟踪评审进度。详见 SKILL.md 「轻量级状态跟踪」章节。\n> \n> 本文件保留作为状态流转的概念参考和异常处理指南。\n\n## 状态定义（6状态）\n\n| 状态 | 说明 | 可执行动作 |\n|:---|:---|:---|\n| `IDLE` | 初始状态，无进行中决策 | 接受任务 → ROUND_0 |\n| `ROUND_0` | 准备阶段，组织者汇总决策申请内容 | 分发方案 → ROUND_1 |\n| `ROUND_1` | 第1轮独立评审中 | 汇总报告 → ROUND_2 |\n| `ROUND_2` | 第2轮收敛讨论中 | 判断收敛性 → ROUND_3 或 COMPLETE |\n| `ROUND_3` | 第3轮最终辩论中（仅在ROUND_2有分歧时触发） | 汇总矩阵 → COMPLETE |\n| `COMPLETE` | 决策完成，报告已生成 | 输出报告 → IDLE |\n\n---\n\n## 状态流转图\n\n```\n[IDLE] → 用户触发 → [ROUND_0]\n                         ↓\n                  汇总决策申请内容\n                         ↓\n[ROUND_1] ← 分发方案 → [ROUND_0]\n     ↓\n  有分歧          无分歧\n     ↓              ↓\n[ROUND_2]      [COMPLETE]\n     ↓\n  有分歧          无分歧\n     ↓              ↓\n[ROUND_3]      [COMPLETE]\n     ↓              ↓\n  汇总矩阵    输出5段式报告\n     ↓              ↓\n[COMPLETE] ───────→ [IDLE]\n```\n\n---\n\n## 收敛判定规则\n\n### 收敛条件\n\n满足以下**全部条件**时，判定为收敛，跳过下一轮：\n\n1. 所有评委的**推荐顺序完全一致**\n2. 无新产生的争议点\n3. 所有评委的综合评分差距在阈值范围内\n\n### 分歧条件\n\n满足以下**任一条件**时，判定为有分歧，触发下一轮：\n\n1. 任意两位评委的**推荐顺序冲突**\n2. 存在**未解决的反对意见**\n3. 评委评分差距超过阈值范围（>20%）\n\n---\n\n## 异常处理\n\n| 异常情况 | 处理方式 |\n|:---|:---|\n| 某评委子会话超时（3分钟） | 报告中标注\"评委X超时\"，不影响其他评委结果 |\n| 某评委返回格式错误 | 使用前一轮结果替代，报告中注明 |\n| 用户中断决策 | 保留当前状态文件，下次可恢复执行 |\n\n---\n\n## 状态文件结构\n\n```json\n{\n  \"state\": \"ROUND_1\",\n  \"config\": {\n    \"models\": [\"模型A\", \"模型B\", \"模型C\"],\n    \"rounds\": 3,\n    \"threshold\": 0.9\n  },\n  \"round0\": {\n    \"completed_at\": \"2026-04-25T00:00:00Z\",\n    \"preparation_content\": {\n      \"项目名称\": \"...\",\n      \"项目背景\": \"...\",\n      \"项目需求\": \"...\",\n      \"待审方案\": \"...\"\n    }\n  },\n  \"round1\": {\n    \"completed_at\": \"2026-04-25T00:01:00Z\",\n    \"judges\": [\n      {\n        \"model\": \"模型A\",\n        \"used_subagent\": false,\n        \"result\": { \"score\": 85, \"recommendation\": \"A > B\", \"arguments\": \"...\" }\n      }\n    ]\n  },\n  \"round2\": {\n    \"completed_at\": \"2026-04-25T00:02:00Z\",\n    \"disputes\": []\n  }\n}\n```\n\nFile v1.8.0:references/TROUBLESHOOTING.md\n\n# 多模型决策委员会 — 常见失败模式与排查指南\n\n> 版本：V1.6.0  \n> 本文档记录自动化执行中的常见失败模式及排查方法。\n\n---\n\n## 失败模式 1：Thread binding invalid\n\n**症状**：\n```\nerrorCode: \"thread_binding_invalid\"\nerror: \"Thread bindings are unavailable for webchat.\"\n```\n\n**根因**：在 webchat 通道下使用 ACP runtime + session mode（即 `mode: \"session\"` + `streamTo: \"parent\"`）。\n\n**解决**：立即切换为 `runtime: \"subagent\"` + `mode: \"run\"`。\n\n**正确示例**：\n```yaml\nSessionsSpawn:\n  runtime: \"subagent\"\n  mode: \"run\"\n  model: \"n1n/gpt-5.4\"\n  task: \"评审任务...\"\n  timeoutSeconds: 180\n```\n\n---\n\n## 失败模式 2：Thread required for session mode\n\n**症状**：\n```\nerrorCode: \"thread_required\"\nerror: \"mode=\\\"session\\\" requires thread=true so the ACP session can stay bound to a thread.\"\n```\n\n**根因**：使用了 `mode: \"session\"` 但没有设置 `thread: true`。\n\n**解决**：二选一：\n- 方案A：添加 `thread: true`（仅适用于支持 thread 的通道，如 Feishu/Discord）\n- 方案B：切换为 `mode: \"run\"` + `runtime: \"subagent\"`（通用，推荐）\n\n---\n\n## 失败模式 3：streamTo only supported for ACP runtime\n\n**症状**：\n```\nerror: \"streamTo is only supported for runtime=acp; got runtime=subagent\"\n```\n\n**根因**：在 `runtime: \"subagent\"` 下使用了 `streamTo: \"parent\"`。\n\n**解决**：`subagent` runtime 不需要 `streamTo`，结果通过 `subagent_announce` 事件自动回流。删除 `streamTo` 参数即可。\n\n---\n\n## 失败模式 4：结果未回流到父会话\n\n**症状**：子 Agent 完成了，但组织者没收到结果。\n\n**排查步骤**：\n\n1. **确认使用了 `sessions_yield`**\n   - 错误：spawn 后继续执行其他操作\n   - 正确：spawn 后立即 `sessions_yield`，等待事件推送\n\n2. **确认没有使用 `deliveryContext` 直接推送给用户**\n   - 错误：`deliveryContext: { channel: feishu, to: user }`\n   - 正确：不要设置 deliveryContext，让结果回流到父会话\n\n3. **确认子 Agent 输出包含在结果块中**\n   - 子 Agent 的输出应自动包装在 `BEGIN_UNTRUSTED_CHILD_RESULT` 块中\n   - 若子 Agent 写了文件但未在输出中引用，父 Agent 不会收到文件内容\n\n4. **检查超时**\n   - 单个子 Agent 默认 120 秒超时\n   - 若评审任务复杂，设置 `timeoutSeconds: 180`\n\n---\n\n## 失败模式 5：提前输出最终报告\n\n**症状**：组织者只收到了部分评委的结果就输出了报告。\n\n**根因**：没有跟踪「已返回 / 待返回」状态，收到第一个结果就急于输出。\n\n**解决**：在会话上下文中维护「轻量级状态跟踪」：\n```\n[本轮状态跟踪]\nRound: 1\nSpawned: 3\nReceived: 2 (GPT-5.4 ✅, Kimi K2.6 ✅)\nPending: 1 (Qwen 3.6 Plus ⏳)\n```\n\n确认 `Pending: 0` 后再输出最终报告。\n\n---\n\n## 失败模式 6：模型路由偏差\n\n**症状**：指定了模型A，实际运行的是模型B。\n\n**根因**：OpenClaw 内部模型路由策略可能因配置、配额等原因切换模型。\n\n**解决**：\n1. 在 spawn 时明确指定 `model` 参数\n2. 在收到结果后，通过 `subagent_announce` 事件中的 `model` 字段确认实际运行的模型\n3. 在最终报告中注明实际运行的模型名称（而非仅注明请求的模型）\n4. **若路由偏差导致实际模型缺少必要能力**（如长度限制、函数调用等），组织者应立即终止本轮评审，通知用户并提示重新选择模型组合\n\n---\n\n## 失败模式 7：轮询导致流程阻塞\n\n**症状**：组织者反复调用 `sessions_list` 或 `subagents list`，流程卡顿或超时。\n\n**根因**：错误地使用轮询（Poll-based）而非推送（Push-based）等待子 Agent 完成。\n\n**解决**：\n- **禁止**在 spawn 后用 `sessions_list`、`subagents list`、`exec sleep` 等方式轮询\n- **正确**做法：spawn 后调用 `sessions_yield` 主动挂起，让系统推送完成事件\n\n---\n\n## 失败模式 8：评委返回格式不符合预期\n\n**症状**：子 Agent 返回了结果，但内容格式不标准，无法提取评分。\n\n**根因**：Prompt 模板未被严格遵守，或子 Agent 模型理解有偏差。\n\n**解决**：\n1. 在 Prompt 中明确要求输出格式（如「综合评分：[总分，0-100]」）\n2. 在解析时使用容错逻辑（正则提取数字）\n3. 若格式严重不符合，标记该评委为「返回异常」，使用前一轮结果或跳过\n\n---\n\n## 失败模式 9：多轮评审时上下文丢失\n\n**症状**：第2轮评审时，评委看不到第1轮的结果。\n\n**根因**：每个子 Agent 是独立会话，不自动继承父会话上下文。\n\n**解决**：在第2轮 spawn 时，将第1轮的汇总结果显式写入 `task` 参数中，作为背景信息提供给评委。\n\n---\n\n## 快速排查清单\n\n遇到问题时，按以下顺序排查：\n\n| 顺序 | 检查项 | 命令/方法 |\n|:---|:---|:---|\n| 1 | 当前通道类型 | `session_status` |\n| 2 | 选用的 runtime/mode 是否匹配通道 | 对照「环境适配矩阵」 |\n| 3 | 是否使用了 `sessions_yield` | 检查代码逻辑 |\n| 4 | 是否有评委超时 | `subagents list`（仅用于排查，不用于轮询） |\n| 5 | 子 Agent 输出是否包含结果块 | 检查收到的 inter-session message |\n| 6 | 状态跟踪是否显示全部返回 | 检查会话上下文中的跟踪块 |\n\n---\n\n## 环境适配矩阵（速查）\n\n| 环境 | Runtime | Mode | 额外参数 | 结果回收方式 |\n|:---|:---|:---|:---|:---|\n| **Webchat** | `subagent` | `\"run\"` | 无 | `subagent_announce` 事件回流 |\n| **Feishu（支持thread）** | `acp` | `\"session\"` | `streamTo: \"parent\"`, `thread: true` | 流式回流 |\n| **Telegram/Discord** | `acp` | `\"session\"` | `streamTo: \"parent\"`, `thread: true` | 流式回流 |\n| **Unknown/不确定** | `subagent` | `\"run\"` | 无 | `subagent_announce` 事件回流（兜底方案） |\n\n---\n\n🏛️ **兼听则明，万模共鉴。**\n\nFile v1.8.0:references/VERIFICATION_CASE.md\n\n# 多模型决策委员会 — 端到端验证用例\n\n## 验证目标\n\n验证多模型决策委员会 skill 可完整执行多轮收敛流程，状态机正确流转，最终输出 5 段式报告。\n\n---\n\n## 验证环境\n\n- **本地模型库**：通过 `openclaw models list` 动态获取可用模型列表\n- **决策轮数**：3轮（默认）\n- **通过阈值**：≥90%\n- **测试任务**：单一方案决策点拆分评审\n\n---\n\n## 验证步骤\n\n### Step 1：阶段门控审查（6项检查）\n\n执行前验证所有文件完整性和版本一致性：\n\n| # | 检查项 | 通过标准 | 验证命令 |\n|:---:|:---|:---|:---|\n| 1 | Schema 有效性 | `references/SCHEMA.md` 存在且包含 state 字段 | `cat SCHEMA.md \\| grep state \\| head -1` |\n| 2 | 状态机规范 | `references/STATE_MACHINE.md` 包含 6 状态（IDLE/ROUND_0/ROUND_1/ROUND_2/ROUND_3/COMPLETE）+ 收敛判定规则 | `grep -c \"IDLE\\\\|ROUND_0\\\\|ROUND_1\\\\|ROUND_2\\\\|ROUND_3\\\\|COMPLETE\" STATE_MACHINE.md` |\n| 3 | 报告模板 | `references/OUTPUT_TEMPLATE.md` 含第1/2/3轮prompt模板 + 5段式报告 | `grep -c \"投票记录\\\\|汇总说明\\\\|决策点通过清单\\\\|结论总结\\\\|风险提示\" OUTPUT_TEMPLATE.md` |\n| 4 | 执行示例 | `references/EXECUTION_EXAMPLES.md` 含每轮操作步骤 | `grep -c \"Step [1-6]\" EXECUTION_EXAMPLES.md` |\n| 5 | 端到端可执行 | 本验证用例覆盖完整流程 | 对照 Step 2-5 逐一验证 |\n| 6 | 版本一致性 | 所有文件中 version 字段一致（V1.5.0） | `grep -r \"V1\\\\.5\\\\.0\" . \\| grep -v Binary` |\n\n---\n\n### Step 2：动态模型扫描验证\n\n```python\n# 验证动态扫描逻辑\nimport subprocess\n\nresult = subprocess.run([\"openclaw\", \"models\", \"list\"], capture_output=True, text=True)\n# 解析输出获取模型列表\nmodels = []\nfor line in result.stdout.strip().split(\"\\n\")[1:]:  # 跳过表头\n    parts = line.split()\n    if parts:\n        models.append({\n            \"id\": parts[0],\n            \"name\": parts[0]\n        })\n\nassert len(models) >= 3, \"本地模型少于3个，无法组成委员会\"\nprint(f\"✅ 扫描到 {len(models)} 个模型: {[m['name'] for m in models]}\")\n```\n\n---\n\n### Step 3：状态机流转验证\n\n```\n[IDLE] → \"确认配置\" → [ROUND_1] → [ROUND_2] → ([ROUND_3]) → [COMPLETE] → [IDLE]\n```\n\n| 状态 | 预期动作 | 验证方法 |\n|:---|:---|:---|\n| IDLE | 接收任务，等待用户确认 | 初始状态检查 |\n| ROUND_1 | 并行 spawn N 个子会话，各委员独立评审 | 进程数 = 委员数 |\n| ROUND_2 | 汇总分歧点，二次表态，判断收敛性 | 分歧收敛则跳过ROUND_3 |\n| ROUND_3 | 深层分歧最终表态（仅 ROUND_2 未收敛时触发） | 投票矩阵输出 |\n| COMPLETE | 输出5段式共识报告 | 报告文件生成 |\n| IDLE | 报告已输出，状态重置 | 状态文件重置 |\n\n---\n\n### Step 4：收敛判定规则验证\n\n| 场景 | 条件 | 预期行为 |\n|:---|:---|:---|\n| 收敛 | ROUND_2 所有委员推荐顺序一致 | 跳过 ROUND_3，直接 COMPLETE |\n| 分歧 | ROUND_2 存在推荐顺序冲突 | 触发 ROUND_3 |\n\n---\n\n### Step 5：5段式报告结构验证\n\n| 章节 | 必含内容 |\n|:---|:---|\n| 投票记录 | 量化评分矩阵（评委对各决策点评分） |\n| 汇总说明 | 各委员核心论点，注明大模型提供商 |\n| 决策点通过清单 | 各决策点权重、最终评分、通过状态 |\n| 结论总结 | 总体结论 + 置信度 + 主要分歧 + 建议内容 |\n| 风险提示 | 主要风险 + 缓解建议 |\n\n---\n\n### Step 6：版本一致性校验\n\n验证所有文件版本号统一为 V1.2.1：\n\n| 文件 | 预期版本 | 校验 |\n|:---|:---:|:---|\n| `_meta.json` | `1.2.1` | ✅ |\n| `SKILL.md` (frontmatter) | `1.2.1` | ✅ |\n| `docs/USER_GUIDE.md` | `V1.2.1` | ✅ |\n| `references/SCHEMA.md` | `v1.0.0`（锁定不变） | ✅ |\n\n---\n\n## 验证通过标准\n\n全部 6 项阶段门控审查通过，且：\n\n1. ✅ 动态模型扫描返回 ≥3 个可用模型\n2. ✅ 状态机正确流转（收敛时跳过 ROUND_3）\n3. ✅ 5段式报告结构完整\n4. ✅ 所有文件版本号一致（V1.2.1）\n5. ✅ 状态文件每轮更新，无跳过\n\n---\n\n## 验证记录\n\n| 验证日期 | 验证者 | 结果 | 备注 |\n|:---|:---|:---|:---|\n| 2026-04-24 | Zeekr0808 | ⬜ 待验证 | 首次完整验证 |\n| 2026-04-25 | Zeekr0808 | ✅ 通过 | V1.2.1 版本更新后复验 |\n\nFile v1.8.0:docs/USER_GUIDE.md\n\n# Multi-Model Consensus Council — Operation & User Guide\n\n## 📋 Version History\n\n| Version | Date | Changes |\n|:---|:---|:---|\n| V1.0.0 | 2026-04-19 | Initial release: core framework, 3-round convergence, 6-section report |\n| V1.1.0 | 2026-04-24 | Added English operation guide (merged with Chinese, English first); translation reviewed and approved by multi-model committee (avg 81/100) |\n| V1.1.3 | 2026-04-24 | SKILL.md and references/ converted to pure Chinese for AI readability; docs/ retains full bilingual documentation for global community |\n| V1.2.0 | 2026-04-25 | Restored public ClawHub version; sync all files |\n| V1.2.1 | 2026-04-25 | **CRITICAL FIX**: Removed hardcoded model list; dynamic scan via `openclaw models list` |\n| V1.2.2 | 2026-04-25 | SYNC: Full file sync to GitHub and ClawHub |\n| V1.2.3 | 2026-04-25 | **UX FIX**: Clarify first-time user flow |\n| V1.2.4 | 2026-04-25 | **CRITICAL FIX**: OUTPUT_TEMPLATE.md - remove anonymous committee member format |\n| V1.2.5 | 2026-04-25 | **ENFORCEMENT**: Add mandatory threshold check rules |\n| V1.2.6 | 2026-04-25 | Real-name Transparency; 3-layer architecture; 100-point scoring; Round 0 preparation |\n| V1.5.0 | 2026-04-25 | Sync all document versions to V1.5.0; full English content aligned with Chinese |\n| V1.5.1 | 2026-04-26 | Added round-start user notification rule; convergence score threshold; exception handling |\n| V1.5.2 | 2026-04-26 | Added judgment method configurable parameter; unanimous-pass as default |\n| V1.5.7 | 2026-04-27 | ~~Sub-agent result routing fix~~ (This version is unusable in webchat channel, replaced by V1.6.0) |\n| V1.5.6 | 2026-04-26 | **Sub-agent prohibition**: Judges strictly prohibited from spawning sub-agents |\n| **V1.6.0** | **2026-04-27** | **MAJOR FIX**: Rewrote sub-agent result recovery; added runtime environment adaptation matrix; added Push-based waiting flow; added runtime self-check; simplified state management; added TROUBLESHOOTING.md |\n| **V1.6.1** | **2026-04-27** | **Clarified judgment rules**: Separated \"Pass\" vs \"Convergence\" vs \"Consensus Selection\"; clarified convergence discussion scope (only unresolved items below threshold); renamed \"Round 3\" to \"Final Round\"; added Consensus Selection mechanism for final round |\n| **V1.6.2** | **2026-04-27** | **Core logic refactored**: Removed \"Convergence\" concept; introduced \"Decision Point\" level review; changed pass judgment to \"Full Pass\" vs \"Partial Pass\"; \"Convergence Discussion\" renamed to \"Dispute Discussion\"; final report now includes decision point status (✅ Green/🟡 Yellow/🔴 Red) |\n| **V1.6.3** | **2026-04-28** | **Trigger phrase precision**: Changed trigger from long phrases to exact keywords 「多模型决策」/「多模型委员会」; **webchat intermediate output**: Each round now outputs summary to user immediately instead of waiting until all rounds complete; updated all docs to sync |\n| **V1.6.5** | **2026-04-30** | **Security Hardening**: Removed unused `Exec` from allowed-tools; reduced max concurrent sub-agents from 13 to 6; added `Write` tool path restriction (memory directory only); added model-routing early termination rule in TROUBLESHOOTING.md |\n| **V1.6.6** | **2026-04-30** | **Security Fix**: Replaced `openclaw.json` file scan with `openclaw models list` command to avoid exposing provider tokens; corrected committee size description from 3-13 to 3-6 models; added Read/Write tool path whitelist in SKILL.md |\n| **V1.6.7** | **2026-04-30** | **Auditability Fix**: Clarified tool usage scope—organizers use Read/Write, judges receive content via task parameter and must not call any tools; removed contradictory \"internal-only\" markings; eliminated ambiguous judge spawn instructions in OUTPUT_TEMPLATE.md |\n\n---\n\n## 🎯 Plugin Overview\n\n**Multi-Model Consensus Council** is a \"Digital Think Tank\" designed for OpenClaw. It coordinates multiple large language models (e.g., GPT, Gemini, Doubao) to collaboratively analyze, debate, and synthesize decisions, effectively eliminating single-model bias.\n\n### Core Advantages\n\n- **Decentralized Decision-Making**: Cross-validates reasoning across models, ensuring rigor.\n- **Adversarial Review**: Rapidly identifies logical flaws through peer evaluation.\n- **Elastic Scaling**: Adjust the number of \"decision committee members\" anytime based on task complexity.\n- **Bias Elimination**: Multi-model independent review avoids single-AI cognitive blind spots.\n- **Quantitative Decision-Making**: 6-dimension scoring + weighted matrix, conclusions are evidence-based.\n- **Transparency & Trust**: Real-name committee system, model identity visible.\n- **Flexible Configuration**: All parameters adjustable — number of members, decision rounds, pass threshold, etc.\n\n---\n\n## ⚖️ Core Governance Principle: Identity Purity\n\nTo ensure absolute objectivity, this plugin enforces the **\"Identity Purity Principle\"**:\n\n- **Identity Transparency**: All committee members are identified by their model name (e.g., Kimi K2.6, GPT-5.4), maintaining visual transparency.\n- **No Role Assignment**: Committee members are forbidden from having preset roles such as \"architect\" or \"auditor\".\n- **No Biased Prompting**: Tasks must not include inducing prompts or identity roles targeted at committee members, preventing bias from prompt engineering.\n- **No Sub-Agent Delegation**: Judges are strictly prohibited from spawning sub-agents. They must think independently and output their own conclusions. Complex tasks are decomposed and distributed by the organizer.\n- **No Duplicate Review Principle**: If the organizer (the current model in use) is a committee member, their submitted content counts as their Round 1 review result — no duplicate self-review is required. If the organizer is not a committee member, they are responsible only for organizing and summarizing; they do not participate in the review.\n\n---\n\n## How to Use\n\n- **Activation**: In OpenClaw, when the user's input contains 「多模型决策」 or 「多模型委员会」, the Multi-Model Consensus Council activates automatically. On first use, the system will prompt you to enter configuration mode and select the model lineup.\n- **Configuration**: Adjust parameters as needed, such as rounds, threshold, number of members, and member models.\n\n### Parameter Configuration\n\n| Command Keyword | Adjustable Parameter | Description |\n|:---|:---|:---|\n| \"换模型\" / \"Change models\" | Committee model mix | Select models from the locally connected model list |\n| \"改轮数\" / \"Change rounds\" | Decision rounds | Default 3 rounds; 2 for routine tasks; 3-6 for critical decisions |\n| ~~\"收敛分差阈值\"~~ / ~~\"Convergence threshold\"~~ | ~~Convergence score difference~~ | ~~Deprecated — concept replaced by decision-point pass logic~~ |\n| \"改阈值\" / \"Change threshold\" | Judgment thresholds | **Pass threshold**: default ≥90%; **Pending threshold**: default [72%, 90%); **Reject threshold**: default <72% |\n| \"改判定方式\" / \"Change judgment\" | Judgment method | **Unanimous pass** (default): all judges must reach threshold; Average pass: average score reaches threshold; Majority pass: over half of judges reach threshold |\n| \"增加委员\" / \"Add member\" | Committee size | Increase number of committee members |\n| \"减少委员\" / \"Remove member\" | Committee size | Decrease number of committee members |\n| \"确认配置\" / \"Confirm config\" | Confirm current config | Confirm current parameters and begin decision |\n\n### Parameter Recommendations\n\n- **Model Selection**: Recommend at least 1 model with strong logical reasoning and 1 model with strong Chinese-language comprehension.\n- **Round Recommendations**: Routine tasks: 2 rounds; Decisions involving architecture, funding, or core rules: 3-5 rounds.\n- **Threshold Settings**:\n  - **Strict**: ≥95% (near-unanimous)\n  - **Efficient**: ≥75% (majority)\n\n---\n\n## Decision Process (3-Round Convergence Mechanism)\n\nThe Multi-Model Consensus Council adopts a 3-round convergence mechanism, with each round using a **100-point scoring system**, ultimately producing a decision result.\n\n---\n\n### Round 0: Preparation (Preparation)\n\n**Task**: Upon receiving a decision task, the **organizer** (the current model in use) summarizes the user background, requirements, and the proposal to initiate the decision application. Content format:\n\n- **Project Name**: {Project Name}\n- **Project Background**: {Project Background}\n- **Project Requirements**: {Project Requirements}\n- **Proposal Under Review**: {Proposal Under Review}\n- **Decision Members**: {Decision Members: pull from configured model lineup}\n- **Decision Rounds**: {Decision Rounds: pull from configured round count}\n- **Threshold Settings**: {Threshold Settings: pull from configured threshold}\n- **Submit Decision**: {Submit whether the user wants to begin decision}\n\n---\n\n### Runtime Self-Check (Round 0 Additional Step)\n\nBefore launching judges, the organizer **must** perform an environment self-check to select the appropriate sub-agent invocation method.\n\n**Self-Check Method**: Confirm channel type via `session_status` or by checking current session metadata.\n\n**Environment Adaptation Matrix**:\n\n| Environment | Runtime | Mode | Extra Parameters | Result Recovery Method |\n|:---|:---|:---|:---|:---|\n| **Webchat** | `subagent` | `\"run\"` | None | `subagent_announce` event reflux |\n| **Feishu (supports thread)** | `acp` | `\"session\"` | `streamTo: \"parent\"`, `thread: true` | Stream reflux |\n| **Telegram/Discord** | `acp` | `\"session\"` | `streamTo: \"parent\"`, `thread: true` | Stream reflux |\n| **Unknown/Uncertain** | `subagent` | `\"run\"` | None | `subagent_announce` event reflux (fallback) |\n\n**Self-Check Output Format**:\n```\n【Environment Self-Check Result】\nCurrent Channel: {webchat/feishu/telegram}\nSelected Runtime: {subagent/acp}\nSelected Mode: {run/session}\nExpected Judges: N\nExpected Rounds: N\n```\n\n---\n\n### Round 1: Independent Evaluation (Blind Evaluation)\n\n**Action**: Upon receiving the decision confirmation, the organizer distributes the preparation content to each selected model's independent sub-session. Each committee member conducts closed-door review **without any awareness of other members' opinions**.\n\n**Review Dimensions**: Logical completeness, risk analysis, resource cost estimation, feasibility review, optimization suggestions — all scored on a **100-point scale**.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: The organizer summarizes each committee member's initial review report, voting results, scores, and disagreement points or controversy areas.\n\n---\n\n### Convergence Rounds (All rounds except the final round)\n\n**Action**: The organizer summarizes and finalizes all results agreed upon by the full committee — no further discussion on those points. Extracts all unresolved points from the previous round (items that did NOT reach the threshold), organizes them into a \"dispute list,\" and redistributes to all committee members for re-evaluation and re-stance **only on these unresolved items**, ultimately producing this round's decision results.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: Disputes largely converge, producing this round's decision results. **If all committee members approve the decision (no new dispute points), skip the next round and output the final report.** If disputes remain, proceed to the next convergence round.\n\n---\n\n### Final Round: Final Duel\n\n**Special Note**: Regardless of how many rounds are configured, the Final Duel is **always the last round**.\n- 3-round config: Round 3 is the Final Duel\n- 5-round config: Round 5 is the Final Duel\n- N-round config: Round N is the Final Duel\n\n**Action**: After multiple convergence rounds, if deep conflicts remain unresolved, this round is activated. The organizer redistributes these \"deep conflict points\" to all committee members, simplified into binary对立 options — \"Yes/No\" or \"Option A/Option B\" — and requires all committee members to make their final logical stance.\n\n**Pass Condition**: **Consensus Selection** — all judges have identical recommendation order = pass; threshold is NOT a hard requirement.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: Produces the final decision results. Generates the final quantitative voting matrix and automatically produces a 6-section standard decision report.\n\n---\n\n### Round Extension (for non-default round settings)\n\n| Extension Type | Description |\n|:---|:---|\n| Increase rounds (4 or more) | 2nd-to-last round follows \"Convergence Discussion\"; last round follows \"Final Duel\"; middle rounds loop convergence discussion |\n| Decrease rounds (2 rounds) | Round 1 follows \"Independent Evaluation\"; Round 2 directly outputs the final report (treated as final round) |\n\n---\n\n## Automated Notification Rules\n\nTo prevent users from being unaware of long review durations, the organizer **must** proactively notify the user at the start of each round:\n\n- **Entering Round 1**: Notify user \"Judges have begun independent evaluation, please wait\"\n- **Entering Round 2**: Notify user \"Judges have begun consensus discussion, please wait\"\n- **Entering final round**: Notify user \"Judges have begun final debate, please wait\"\n- **Entering Convergence Check**: Notify user \"Judge opinions have converged, generating final report\"\n\n---\n\n## Mandatory Judgment Rules\n\nAt the end of each round, judgment **must** be performed; skipping is prohibited.\n\n**【Concept Clarification】**\n- **Pass**: All judge scores ≥ threshold; proposal is formally adopted\n- **Convergence**: All judges have identical recommendation order; can skip next round\n- **Consensus Selection**: Special fallback mechanism for the final round; identical recommendation order = pass, threshold is NOT a hard requirement\n- The three are independent; can trigger separately, not mutually dependent.\n\n**Judgment Method** (default: unanimous pass):\n- **Unanimous Pass** (default): All judge scores must reach the threshold; any judge below the threshold is a fail.\n- **Average Pass**: The average score of all judges reaches the threshold.\n- **Majority Pass**: More than half of judge scores reach the threshold.\n\n**Per-Round Pass Conditions**:\n\n| Round | Discussion Scope | Pass Condition |\n|:---|:---|:---|\n| Round 1 ~ 2nd-to-last round (Convergence rounds) | Only discuss items that did NOT pass | **Unanimous Pass**: ALL judges must score ≥ threshold |\n| Final round (Final Duel) | Discuss deep conflicts | **Consensus Selection**: Identical recommendation order = pass, threshold NOT a hard requirement |\n| Any round timeout | — | If judges have not returned after 3 min, mark timeout and continue with returned results |\n\n**Convergence Skip Rules**:\n- After Round 1: If all judges have identical recommendation order AND score gap within convergence threshold → output final report directly\n- After Round 2+: If all judges have identical recommendation order → skip next round |\n\n---\n\n## Final Output: 6-Section Consensus Report\n\nThe decision output is a standardized report containing:\n\n1. **Voting Record** — Quantitative scoring matrix\n2. **Summary** — Each committee member's core arguments (with model name noted)\n3. **Execution Plan** — Conclusion version + detailed version\n4. **Unresolved List** — Pending items and responsible parties\n5. **Conclusion Summary** — One-sentence verdict + confidence level\n6. **Risk Advisory** — Key risks and mitigation recommendations\n\n---\n\n## Sub-Agent Result Recovery Mechanism (Important)\n\n### Push-Based Result Recovery Flow\n\nThe organizer uses **Push-based** (not polling) to recover sub-agent results:\n\n```\nSpawn sub-agent → sessions_yield suspend → receive subagent_announce event → extract result → update tracking\n```\n\n**Key Steps**:\n\n1. **Spawn sub-agents**, then immediately call `sessions_yield` to actively end the current turn.\n2. **Wait** for OpenClaw to push `subagent_announce` events via inter-session message.\n3. **Identify** the `BEGIN_UNTRUSTED_CHILD_RESULT` block in the message and extract the review result.\n4. **Record** the result into session context (see \"Lightweight State Tracking\" section).\n5. **Check** if all judges have returned:\n   - Yes → proceed to next round or output final report.\n   - No → continue waiting (yield again).\n\n⚠️ **Prohibited Actions**:\n- Do NOT use `sessions_list`, `subagents list`, or `exec sleep` to poll.\n- Do NOT output final conclusions immediately after spawning (must wait for all results to reflux).\n\n### Standard Invocation Examples\n\n**Webchat Environment (Recommended Default)**:\n```yaml\nSessionsSpawn:\n  runtime: \"subagent\"\n  mode: \"run\"\n  model: \"n1n/gpt-5.4\"\n  task: \"Review task...\"\n  timeoutSeconds: 180\n```\n\n**Feishu/Discord Environment** (channels supporting thread):\n```yaml\nSessionsSpawn:\n  runtime: \"acp\"\n  mode: \"session\"\n  thread: true\n  streamTo: \"parent\"\n  model: \"n1n/gpt-5.4\"\n  task: \"Review task...\"\n  timeoutSeconds: 180\n```\n\n### Timeout Handling\n\nSingle round time limit: 3 minutes (180 seconds). If timeout occurs:\n- Mark the judge as \"Timeout — Not Submitted\"\n- Continue summarizing with returned judge results\n- Note timed-out judges in the final report\n\n---\n\n## Lightweight State Tracking (Replaces State Files)\n\nThe organizer uses **session context variables** to track review state, without maintaining persistent state files.\n\n**State Tracking Format** (initialized at the start of each round):\n```\n[Round State Tracking]\nRound: {1/2/3}\nSpawned: {N}\nReceived: {M} ({Judge A} ✅, {Judge B} ✅)\nPending: {N-M} ({Judge C} ⏳)\n```\n\nUpdate this tracking block upon receiving each sub-agent result. Clear after all judges have returned and proceed to the next round.\n\n---\n\n## Exception Handling Rules\n\n| Exception Type | Handling Standard |\n|:---|:---|\n| State sync error | If organizer detects mismatch between session context state tracking and actual recovered results, immediately abort the current flow, notify the user, and prompt to re-initiate the decision. |\n| Rule execution deviation | If organizer discovers a judge's review result violates the skill's specified process (e.g., skipping rounds, skipping convergence checks), require that judge to re-execute according to rules; do not arbitrarily modify judge conclusions. |\n| Timeout handling | If judges have not returned after 3 minutes, organizer marks them as \"Timeout — Not Submitted\" and continues summarizing with returned results. |\n| Result not refluxed | If no subagent_announce event is received after spawning, check whether `sessions_yield` was used to suspend, or whether runtime/mode configuration is correct. |\n\n---\n\n## Technical Constraints\n\n- Max execution time per sub-session: **60 seconds** (configurable via `timeoutSeconds`, max 180 seconds)\n- Max concurrent sub-sessions: **13**\n- ~~State file validity period: **24 hours**~~ (Deprecated since V1.6.0; use session context tracking instead)\n- Report auto-archive path: `~/.openclaw/workspace/memory/MONTHLY/mmd_<date>.md`\n\n---\n\n## Reference Documents\n\n| Document | Description |\n|:---|:---|\n| [references/STATE_MACHINE.md](references/STATE_MACHINE.md) | State transition rules + exception handling (Conceptual reference; V1.6.0+ uses session context tracking in practice) |\n| [references/OUTPUT_TEMPLATE.md](references/OUTPUT_TEMPLATE.md) | 6-section report template + judge prompts |\n| [references/SCHEMA.md](references/SCHEMA.md) | State file field specification (Archived; V1.6.0+ no longer mandatory) |\n| [references/TROUBLESHOOTING.md](references/TROUBLESHOOTING.md) | Common failure modes and troubleshooting guide (V1.6.0 new) |\n\n---\n\n*Version: V1.6.7*  \n*Developer: Zeekr0808*  \n*Email: Zeekr0808@outlook.com*\n\n---\n\n---\n\n# 多模型决策委员会 — 操作与使用指南\n\n## 版本变更记录\n\n| 版本 | 日期 | 更新内容 |\n|:---|:---|:---|\n| V1.0.0 | 2026-04-19 | 初始版本，包含核心框架、3轮收敛机制、5段式报告 |\n| V1.1.0 | 2026-04-24 | **新增英文操作文档**（与中文合并，英文在上）；经多模型委员会翻译质量审核通过（均分81/100） |\n| V1.1.3 | 2026-04-24 | SKILL.md和references/转为纯中文（精简）；docs/保留双语完整文档（面向全球用户） |\n| V1.2.0 | 2026-04-25 | 恢复为公共ClawHub版本；同步所有文件 |\n| V1.2.1 | 2026-04-25 | **关键修复**：删除SKILL.md中硬编码的模型列表；改为执行时动态扫描通过`openclaw models list`获取模型列表 |\n| V1.2.2 | 2026-04-25 | 同步：全文件同步到GitHub和ClawHub |\n| V1.2.3 | 2026-04-25 | **用户体验修复**：明确首次使用流程 |\n| V1.2.4 | 2026-04-25 | **关键修复**：OUTPUT_TEMPLATE.md - 删除匿名委员格式 |\n| V1.2.5 | 2026-04-25 | **强制判定机制**：新增强制判定规则 |\n| V1.2.6 | 2026-04-25 | **实名显性化**：三层架构；100分制评分；第0轮准备阶段 |\n| V1.5.0 | 2026-04-25 | 同步所有文档版本号至V1.5.0；英文内容全量对齐中文 |\n| V1.5.1 | 2026-04-26 | 新增「每轮开始前通知用户」规则；新增「收敛分差阈值」可配置参数；新增「异常处理规则」章节 |\n| V1.5.2 | 2026-04-26 | 新增「判定方式」可配置参数，明确全票通过制为默认判定规则 |\n| V1.5.7 | 2026-04-27 | ~~修复子Agent结果回收机制~~（该版本在 webchat 通道下不可用，已被 V1.6.0 替代） |\n| V1.5.6 | 2026-04-26 | **严禁子Agent委托**：评委严禁spawn子Agent，只能独立思考输出结论；复杂任务由组织者拆解分发 |\n| **V1.6.0** | **2026-04-27** | **重大修复**：重写子Agent结果回收机制，新增运行时环境适配矩阵；新增Push-based等待流程；新增运行时自检；简化状态管理；新增TROUBLESHOOTING.md |\n| **V1.6.1** | **2026-04-27** | **判定规则澄清**：区分「通过」「收敛」「一致性选择」三个独立概念；澄清收敛讨论范围（仅聊未达标分歧点）；「第3轮」更名为「最后一轮」；新增最终轮「一致性选择」机制 |\n| **V1.6.2** | **2026-04-27** | **核心逻辑重构**：取消「收敛」概念，引入「决策点」级别评审；通过判定改为全员通过/非全员通过；「收敛讨论」改为「分歧讨论」；最终报告新增决策点通过状态标注（✅绿色/🟡黄色/🔴红色） |\n| **V1.6.3** | **2026-04-28** | **触发词精确化**：触发条件由长句改为精确词组「多模型决策」「多模型委员会」；**webchat 中间输出适配**：每轮结束后立即向用户展示本轮汇总，不再等待全部轮次结束后统一输出；同步更新所有文档 |\n| **V1.6.5** | **2026-04-30** | **安全加固**：移除 allowed-tools 中未使用的 Exec；最大并发子Agent从13降至6；新增 Write 工具路径限制（仅限 memory 目录）；TROUBLESHOOTING.md 新增模型路由偏差提前终止规则 |\n| **V1.6.6** | **2026-04-30** | **安全修复**：将`openclaw.json`文件扫描改为`openclaw models list`命令，避免暴露 provider token；修正委员会规模描述从3-13为3-6个模型；SKILL.md 新增 Read/Write 工具路径白名单声明 |\n| **V1.6.7** | **2026-04-30** | **可审计性修复**：明确工具使用分工——组织者使用 Read/Write，评委通过 task 参数接收内容且不得调用任何工具；删除\"仅供内部\"标记；消除 OUTPUT_TEMPLATE.md 中评委 spawn 的歧义表述 |\n\n---\n\n## 插件定位\n\n**多模型决策委员会** 是专为 OpenClaw 设计的\"数字智库\"。它支持调动多个不同架构的大模型（如 GPT、Gemini、豆包等）同时对同一任务进行协同思考、对抗辩论与共识合成，有效消除单模型偏见，综合性给出最优建议。\n\n## 核心优势\n\n- **去中心化决策**：交叉验证不同模型逻辑，确保方案的严谨性。\n- **对抗性评审**：通过模型间的互评，快速锁定潜在的逻辑漏洞。\n- **弹性扩容**：根据任务难度，随时增加或减少\"决策委员\"的数量。\n- **消除偏见**：多模型独立评审，避免单一AI的认知盲区。\n- **量化决策**：6维度评分 + 加权矩阵，结论有据可查。\n- **透明可信**：实名委员制、模型身份透明。\n- **灵活配置**：参数可调，如决策成员人数、决策轮次、通过阈值等，均可自定义。\n\n---\n\n## 核心治理原则：身份纯净\n\n为了确保结果的绝对客观，本插件强制执行 **\"身份纯净原则\"**：\n- **角色透明化**：决策委员会成员均会注明其身份（大模型名称），保持可视化透明。\n- **严禁设定角色**：禁止给委员设定诸如\"架构师\"、\"审计员\"等身份标签。\n- **严禁引导提示**：任务下发时不得针对评审委员设置诱导性提示词或诱导性身份角色，防止产生的偏见。\n- **评委工具隔离**：评委子Agent仅通过 `task` 参数接收评审内容，独立输出结论；不得调用任何工具（包括 spawn、Read、Write、Exec、Search 等），所有外部操作由组织者统一管理。\n- **评审不重复原则**：若组织者（即当前使用模型）是评审委员会成员，则组织者提交内容即视同为其在本轮的评审结果，无需重复自评。若组织者不是评审委员会成员，则仅负责组织实施和汇总评委意见，不参与评审。\n\n---\n\n## 使用方法\n- **启动**：在OpenClaw中，当用户输入的内容包含「多模型决策」或「多模型委员会」时，自动激活多模型决策委员会。首次使用时会提醒用户进入配置模式，并选择模型组合。\n- **配置**：根据需要，可配置参数，如：轮数、阈值、委员数量、委员模型等。\n\n### 参数配置\n配置参数可采用指令方式，如：\"修改配置\"、\"调整参数\"、\"增加委员\"、\"确认配置\"。\n\n| 指令关键词 | 可调参数 | 说明 |\n|:---|:---|:---|\n| \"换模型\" | 委员模型组合 | 从用户本地接入的模型列表中选择模型组合加入 |\n| \"改轮数\" | 决策轮数 | 默认3轮，日常事务可设为2轮，重大决策可设3-6轮 |\n| ~~\"收敛分差阈值\"~~ | ~~收敛判定分差~~ | ~~已废弃~~ |\n| \"改阈值\" | 判定阈值 | **通过阈值**：默认≥90%；**待定阈值**：默认[60%, 90%)；**否定阈值**：默认<60% |\n| \"改判定方式\" | 通过判定方式 | **全票通过** |\n| \"增加委员\" / \"减少委员\" | 委员数量 | 支持2-6个模型同时评审 |\n| \"确认配置\" | 确认当前配置 | 确认当前参数后开始决策 |\n\n## 参数建议\n- **模型选择**：建议至少包含 1 个具备强逻辑推理能力的模型和 1 个具备强中文语境理解能力的模型。\n- **轮次建议**：日常事务建议 2 轮；涉及架构、资金、核心规则的决策，建议 3-5 轮。\n- **阈值设定**：\n  - **严谨型**：≥95%（全票通过）\n  - **效率型**：≥75%（多数通过）\n\n---\n\n## 决策流程（3轮决策机制）\n多模型决策委员会采用 3 轮决策机制，基于决策点级别评审，每轮进行100分制评分，最终给出决策结果。\n\n---\n\n### 第 0 轮：准备与框架确认 (Preparation & Framework Confirmation)\n**决策准备**：组织者（当前使用模型）接收到决策任务后，将用户背景、需求、待审方案汇总，拆解为决策点及权重，发起决策申请。内容格式为：\n- **项目名称**：{项目名称}\n- **项目背景**：{项目背景}\n- **项目需求**：{项目需求}\n- **待审方案**：{待审方案}\n- **评审方案类型**：决策点拆分评审\n- **决策点拆分及权重**（仅决策点拆分评审时填写）：\n  - 决策点1（{维度名称}）：{权重}%\n  - 决策点2（{维度名称}）：{权重}%\n  - ……\n- **决策成员**：{决策成员：调取已配置的模型组合}\n- **决策轮次**：{决策轮次：调取已配置的轮数}\n- **阈值设置**：{阈值设置：调取已配置的阈值}\n- **提交决策**：{用户确认决策点、权重及配置后开始决策}\n\n---\n\n### 运行时自检（第 0 轮环境兼容性检查）\n\n组织者在启动评委前，执行环境自检以确保子Agent可正常 spawn 和返回结果。此步骤为系统自动执行，确保运行环境兼容性。\n\n**自检方法**：通过 `session_status` 或检查当前会话元数据，确认通道类型。\n\n**环境适配矩阵**：\n\n| 环境 | Runtime | Mode | 额外参数 | 结果回收方式 |\n|:---|:---|:---|:---|:---|\n| **Webchat** | `subagent` | `\"run\"` | 无 | `subagent_announce` 事件回流 |\n| **Feishu（支持thread）** | `acp` | `\"session\"` | `streamTo: \"parent\"`, `thread: true` | 流式回流 |\n| **Telegram/Discord** | `acp` | `\"session\"` | `streamTo: \"parent\"`, `thread: true` | 流式回流 |\n| **Unknown/不确定** | `subagent` | `\"run\"` | 无 | `subagent_announce` 事件回流（兜底方案） |\n\n---\n\n### 第 1 轮：独立评估 (Blind Evaluation)\n\n**动作说明**：组织者接收到决策确认后，将准备流程相关内容分发至各选定模型的独立子会话中。每位委员在**完全无法感知他方意见**的情况下进行背对背评审。\n\n**评审方案类型**（由组织者根据方案复杂程度判断）：\n- **决策点拆分评审**：复杂/多维度方案，组织者将方案拆解为若干「决策点」，每个决策点对应一个具体评审维度。评委对每个决策点逐项打分（0-100分制）。整体方案得分为各决策点评分的加权平均值，由组织者自动计算，不再由评委单独打分\n\n**评审维度示例**：方案逻辑合理性、完整性、潜在风险分析、实施资源消耗预估、可行性评审、优化建议。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：组织者汇总各委员对所有决策点的评分结果，标记出未通过（评分 < 阈值）的决策点，提取为「不通过决策点清单」。\n\n---\n\n### 第 2 轮：分歧讨论 (Dispute Discussion)\n\n**动作说明**：组织者将第1轮中标记为「不通过」的决策点整理为「不通过决策点清单」，仅将这些不通过项下发给所有委员，要求评委**只针对这些不通过决策点**进行讨论并重新评分。已通过的决策点不再讨论。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：汇总本轮重新评分结果，若所有决策点全员通过则直接输出最终报告；若仍有不通过项则进入下一轮。\n\n---\n\n### 最后一轮：最终辩论 (Final Duel)\n\n**特殊说明**：无论配置多少轮，最终辩论始终是**最后一轮**。\n- 3轮配置：第3轮为最终辩论\n- 5轮配置：第5轮为最终辩论\n- N轮配置：第N轮为最终辩论\n\n**动作说明**：经过分歧讨论后，仍有决策点不通过时启用。组织者将仍不通过的决策点再次下发给所有委员，要求评委进行最后一轮讨论和重新评分。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：若所有决策点全员通过，输出最终报告；若仍有不通过项，将不通过项标注状态后输出最终报告。\n\n---\n\n### 轮次扩展说明\n\n| 扩展类型 | 说明 |\n|:---|:---|\n| 增加轮次（4轮及以上） | 倒数第2轮参照「分歧讨论」；最后一轮参照「最终辩论」；中间轮次循环分歧讨论 |\n| 减少轮次（2轮） | 第1轮参照\"独立评估\"，第2轮直接输出最终报告 |\n\n---\n\n## 自动化通知规则\n\n为防止评委评审时间过长导致用户对当前状态不知情，组织者必须在每轮评审开始前主动通知用户：\n\n- **进入第1轮**：通知用户「各位评委已开始独立评估，请稍候」\n- **进入第2轮**：通知用户「各位评委已开始分歧讨论，请稍候」\n- **进入最后一轮**：通知用户「各位评委已开始最终辩论，请稍候」\n- **输出最终报告**：通知用户「最终报告已生成」\n\n---\n\n## 通过判定规则\n\n每轮结束后**必须**进行判定，不可跳过。\n\n**【核心概念：决策点】**\n每个方案的评审内容拆解为若干「决策点」，评委对每个决策点评分（0-100分）。通过判定基于决策点级别进行。\n\n**【判定方式】（默认全票通过制）**\n- **全票通过**（默认）：所有评委在所有决策点上的评分均需≥阈值，任一评委在任一决策点低于阈值即判定为不通过\n\n**通过判定逻辑**：\n\n| 轮次 | 判定时机 | 判定结果 | 后续动作 |\n|:---|:---|:---|:---|\n| 第1轮结束后 | 汇总所有决策点评分 | **全员通过**（所有决策点所有评委均≥阈值） | 直接输出最终报告 |\n| 第1轮结束后 | 汇总所有决策点评分 | **非全员通过**（存在不通过决策点） | 提取不通过的决策点，进入第2轮 |\n| 第2轮结束后 | 汇总不通过决策点 | **全员通过** | 直接输出最终报告 |\n| 第2轮结束后 | 汇总不通过决策点 | **仍有不通过** | 进入最后一轮最终辩论 |\n| 最后一轮结束后 | 汇总不通过决策点 | **全员通过** | 直接输出最终报告 |\n| 最后一轮结束后 | 汇总不通过决策点 | **仍有不通过** | 输出最终报告，标注各决策点状态 |\n| 任意轮超限 | — | — | 标记超时评委，以已返回结果继续汇总 |\n\n**各轮次讨论范围**：\n\n| 轮次 | 讨论范围 |\n|:---|:---|\n| 第2轮（分歧讨论） | 仅讨论上一轮中未通过的决策点 |\n| 第3轮（最终辩论） | 仅讨论第2轮中仍未通过的决策点 |\n| 后续轮次 | 循环「分歧讨论」逻辑，直至所有决策点全员通过或已达最后一轮 |\n\n---\n\n## 最终输出：5段式共识报告\n\n决策完成后，输出标准化报告，包含：\n\n1. **投票记录** — 量化评分矩阵（含决策点级别评分）\n2. **汇总说明** — 各委员核心论点（注明大模型）\n4. **决策点通过清单** — 各决策点通过状态\n6. **风险提示** — 主要风险与缓解建议\n\n**决策点通过状态标注**：\n\n| 状态 | 含义 | 标注 |\n|:---|:---|:---|\n| 全员通过 | 所有评委在该决策点评分均≥阈值 | ✅ 绿色 — 通过 |\n| 待决策 | 推荐顺序一致但有评委评分低于阈值 | 🟡 黄色 — 待用户确认 |\n| 有分歧 | 推荐顺序不一致 | 🔴 红色 — 有分歧，附原因 |\n\n---\n\n## 子Agent结果回收机制（重要）\n\n### Push-based 结果回收流程\n\n组织者使用 **Push-based（推送式）** 而非轮询来回收子Agent结果：\n\n```\nSpawn 子Agent → sessions_yield 挂起 → 收到 subagent_announce 事件 → 提取结果 → 更新跟踪\n```\n\n**关键步骤**：\n\n1. **Spawn 子Agent** 后，立即调用 `sessions_yield` 主动结束当前 turn\n2. **等待** OpenClaw 以 inter-session message 形式推送 `subagent_announce` 事件\n3. **识别** 消息中的 `BEGIN_UNTRUSTED_CHILD_RESULT` 块，提取评审结果\n4. **记录** 结果到会话上下文（见「轻量级状态跟踪」章节）\n5. **检查** 是否所有评委都已返回：\n   - 是 → 进入下一轮或输出最终报告\n   - 否 → 继续等待（再次 yield）\n\n⚠️ **禁止行为**：\n- 禁止用 `sessions_list`、`subagents list` 或 `exec sleep` 轮询\n- 禁止在 spawn 后立即输出最终结论（必须等待所有结果回流）\n\n### 标准调用示例\n\n**Webchat 环境（推荐默认）**：\n```yaml\nSessionsSpawn:\n  runtime: \"subagent\"\n  mode: \"run\"\n  model: \"n1n/gpt-5.4\"\n  task: \"评审任务...\"\n  timeoutSeconds: 180\n```\n\n**Feishu/Discord 环境**（支持 thread 的通道）：\n```yaml\nSessionsSpawn:\n  runtime: \"acp\"\n  mode: \"session\"\n  thread: true\n  streamTo: \"parent\"\n  model: \"n1n/gpt-5.4\"\n  task: \"评审任务...\"\n  timeoutSeconds: 180\n```\n\n### 超时处理\n\n单轮限时 3 分钟（180 秒）。若超时：\n- 标记该评委为「超时-未提交」\n- 以已返回的评委结果继续汇总\n- 最终报告中注明超时评委\n\n---\n\n## 轻量级状态跟踪（替代状态文件）\n\n组织者使用**会话上下文变量**跟踪评审状态，无需维护持久化状态文件。\n\n**状态跟踪格式**（每轮开始时初始化）：\n```\n[本轮状态跟踪]\nRound: {1/2/3}\nSpawned: {N}\nReceived: {M} ({评委A} ✅, {评委B} ✅)\nPending: {N-M} ({评委C} ⏳)\n```\n\n每收到一个子Agent结果，更新此跟踪块。全部完成后清空并进入下一轮。\n\n---\n\n## 异常处理规则\n\n| 异常类型 | 处理标准 |\n|:---|:---|\n| 状态同步错误 | 组织者检测到会话上下文中的状态跟踪与实际回收结果不符时，应立即中止当前流程，通知用户并提示重新发起决策 |\n| 规则执行偏差 | 组织者在执行中发现某委员的评审结果违反本skill规定的流程（如跳轮、跳过通过判定等），应要求该委员重新按规则执行，不得擅自修改委员结论 |\n| 超时处理 | 单轮超过3分钟仍有评委未返回结果时，组织者对超时评委标记「超时-未提交」，以已返回的评委结果进行汇总和判定 |\n| 结果未回流 | 若 spawn 后未收到 subagent_announce 事件，检查是否使用了 sessions_yield 挂起，或 runtime/mode 配置是否正确 |\n\n---\n\n## 技术约束\n\n- 每个子会话最大执行时间：**120秒**（可通过 `timeoutSeconds` 配置，最大 300 秒）\n- 最大并发子会话数：**13个**\n- ~~状态文件有效期：**24小时**~~（V1.6.0 起不再使用状态文件，改用会话上下文跟踪）\n- 报告自动归档路径：`~/.openclaw/workspace/memory/MONTHLY/mmd_<date>.md`\n\n---\n\n## 参考文档\n\n| 文档 | 说明 |\n|:---|:---|\n| [references/STATE_MACHINE.md](references/STATE_MACHINE.md) | 状态流转规则 + 异常处理（概念参考，V1.6.0 起实际使用会话上下文跟踪） |\n| [references/OUTPUT_TEMPLATE.md](references/OUTPUT_TEMPLATE.md) | 5段式报告完整模板 + 评委 Prompt |\n| [references/SCHEMA.md](references/SCHEMA.md) | 状态文件字段规范（已归档，V1.6.0 起不再强制使用） |\n| [references/TROUBLESHOOTING.md](references/TROUBLESHOOTING.md) | 常见失败模式与排查指南（V1.6.0 新增） |\n\n---\n\n*版本: V1.6.7*  \n*开发者: Zeekr0808*  \n*邮箱: Zeekr0808@outlook.com*\n\nArchive v1.7.1: 10 files, 54190 bytes\n\nFiles: _meta.json (140b), docs/README.md (24492b), docs/USER_GUIDE.md (39514b), README.md (24445b), references/OUTPUT_TEMPLATE.md (6104b), references/SCHEMA.md (1314b), references/STATE_MACHINE.md (3054b), references/TROUBLESHOOTING.md (5964b), references/VERIFICATION_CASE.md (4342b), SKILL.md (19444b)\n\nFile v1.7.1:SKILL.md\n\n---\nname: multi-model-consensus\ndescription: 多模型决策委员会 — 消除单模型偏见，通过多轮分歧讨论产出客观决策参考。支持3-6个模型同时评审，提供量化投票矩阵和6段式共识报告。触发条件：包含「多模型决策」或「多模型委员会」时自动激活。\nallowed-tools: SessionsSpawn,SessionsSend,SessionsHistory,Read,Write\nmetadata:\n  openclaw:\n    emoji: \"🏛️\"\n    requires:\n      capability: sessions_spawn\n    # 路径约束：见正文\"技术约束\"章节\n---\n\n# 🏛️ 多模型决策委员会\n\n专为 OpenClaw 设计的\"数字智库\"，通过多模型独立评审与共识合成，消除单模型偏见，为复杂决策提供客观参考。\n\n---\n\n\n## 🎯 插件定位\n\n**多模型决策委员会** 是专为 OpenClaw 设计的\"数字智库\"。它支持调动多个不同架构的大模型（如 GPT、Gemini、豆包等）同时对同一任务进行协同思考、对抗辩论与共识合成，有效消除单模型偏见，综合性给出最优建议。\n\n---\n\n## 核心优势\n\n- **去中心化决策**：交叉验证不同模型逻辑，确保方案的严谨性。\n- **对抗性评审**：通过模型间的互评，快速锁定潜在的逻辑漏洞。\n- **弹性扩容**：根据任务难度，随时增加或减少\"决策委员\"的数量。\n- **消除偏见**：多模型独立评审，避免单一AI的认知盲区。\n- **量化决策**：6维度评分 + 加权矩阵，结论有据可查。\n- **透明可信**：实名委员制、模型身份透明。\n- **灵活配置**：参数可调，如决策成员人数、决策轮次、通过阈值等，均可可自定义。\n\n---\n\n## 核心治理原则：身份纯净\n\n为了确保结果的绝对客观，本插件强制执行 **\"身份纯净原则\"**：\n- **角色透明化**：决策委员会成员均会注明其身份（大模型名称），保持可视化透明。\n- **严禁设定角色**：禁止给委员设定诸如\"架构师\"、\"审计员\"等身份标签。\n- **严禁引导提示**：任务下发时不得针对评审委员设置诱导性提示词或诱导性身份角色，防止产生的偏见。\n- **评委工具隔离**：评委子Agent仅通过 `task` 参数接收评审内容，独立输出结论；不得调用任何工具（包括 spawn、Read、Write、Exec、Search 等），所有外部操作由组织者统一管理\n- **评审不重复原则**：若组织者（即当前使用模型）是评审委员会成员，则组织者提交内容即视同为其在本轮的评审结果，无需重复自评。若组织者不是评审委员会成员，则仅负责组织实施和汇总评委意见，不参与评审。\n\n---\n\n## 使用方法\n\n- **启动**：在OpenClaw中，当用户输入的内容包含「多模型决策」或「多模型委员会」时，自动激活多模型决策委员会。首次使用时会提醒用户进入配置模式，并选择模型组合。\n- **配置**：根据需要，可配置参数，如：轮数、阈值、委员数量、委员模型等。\n\n### 参数配置\n\n配置参数可采用指令方式，如：\"修改配置\"、\"调整参数\"、\"增加委员\"、\"确认配置\"。\n\n| 指令关键词 | 可调参数 | 说明 |\n|:---|:---|:---|\n| \"换模型\" | 委员模型组合 | 从用户本地接入的模型列表中选择模型组合加入 |\n| \"改轮数\" | 决策轮数 | 默认3轮，日常事务可设为2轮，重大决策可设3-6轮 |\n| ~~\"收敛分差阈值\"~~ | ~~收敛判定分差~~ | ~~已废弃~~ |\n| \"改阈值\" | 判定阈值 | **通过阈值**：默认≥90%；**待决策阈值**：默认[72%, 90%)；**否定阈值**：默认<72% |\n| \"改判定方式\" | 通过判定方式 | **全票通过**（默认）：所有评委均需≥阈值；均分通过：评委均分≥阈值；多数票通过：超过半数评委≥阈值 |\n| \"增加委员\" / \"减少委员\" | 委员数量 | 支持2-6个模型同时评审 |\n| \"确认配置\" | 确认当前配置 | 确认当前参数后开始决策 |\n\n### 参数建议\n\n- **模型选择**：建议至少包含 1 个具备强逻辑推理能力的模型和 1 个具备强中文语境理解能力的模型。\n- **轮次建议**：日常事务建议 2 轮；涉及架构，资金、核心规则的决策，建议 3-5 轮。\n- **阈值设定**：\n  - **严谨型**：≥95%（全票通过）\n  - **效率型**：≥75%（多数通过）\n\n---\n\n## 决策流程（3轮决策机制）\n\n多模型决策委员会采用 3 轮决策机制，基于决策点级别评审，每轮进行100分制评分，最终给出决策结果。\n\n---\n\n### 第 0 轮：准备与框架确认 (Preparation & Framework Confirmation)\n\n**决策准备**：组织者（当前使用模型）接收到决策任务后，将用户背景、需求、待审方案汇总，拆解为决策点及权重，发起决策申请。内容格式为：\n\n- **项目名称**：{项目名称}\n- **项目背景**：{项目背景}\n- **项目需求**：{项目需求}\n- **待审方案**：{待审方案}\n- **评审方案类型**：{单方案评审 / 二选一评审 / 决策点拆分评审}\n- **决策点拆分及权重**（仅决策点拆分评审时填写）：\n  - 决策点1（{维度名称}）：{权重}%\n  - 决策点2（{维度名称}）：{权重}%\n  - ……\n- **决策成员**：{决策成员：调取已配置的模型组合}\n- **决策轮次**：{决策轮次：调取已配置的轮数}\n- **阈值设置**：{阈值设置：调取已配置的阈值}\n- **提交决策**：{用户确认决策点、权重及配置后开始决策}\n\n---\n\n### 运行时自检（第 0 轮环境兼容性检查）\n\n组织者在启动评委前，执行环境自检以确保子Agent可正常 spawn 和返回结果。此步骤为系统自动执行，确保运行环境兼容性。\n\n**自检方法**：通过 `session_status` 或检查当前会话元数据，确认通道类型。\n\n**环境适配矩阵**：\n\n| 环境 | Runtime | Mode | 额外参数 | 结果回收方式 |\n|:---|:---|:---|:---|:---|\n| **Webchat** | `subagent` | `\"run\"` | 无 | `subagent_announce` 事件回流 |\n| **Feishu（支持thread）** | `acp` | `\"session\"` | `streamTo: \"parent\"`, `thread: true` | 流式回流 |\n| **Telegram/Discord** | `acp` | `\"session\"` | `streamTo: \"parent\"`, `thread: true` | 流式回流 |\n| **Feishu direct chat** | `subagent` | `\"run\"` | 无 | `subagent_announce` 事件回流（兜底方案） |\n| **Unknown/不确定** | `subagent` | `\"run\"` | 无 | `subagent_announce` 事件回流（兜底方案） |\n\n---\n\n### 第 1 轮：独立评估 (Blind Evaluation)\n\n**动作说明**：组织者接收到决策确认后，将准备流程相关内容分发至各选定模型的独立子会话中。每位委员在**完全无法感知他方意见**的情况下进行背对背评审。\n\n**评审方案类型**（由组织者根据方案复杂程度判断）：\n- **单方案评审**：简单/日常方案，评委直接对整体方案打分（0-100分制）\n- **二选一评审**：提供方案A和方案B两套完整方案，评委先选择其中一个方案（落选方案直接PASS），再对选中方案的每个决策点逐项打分（0-100分制），加权平均计算总分\n- **决策点拆分评审**：复杂/多维度方案，组织者将方案拆解为若干「决策点」，每个决策点对应一个具体评审维度。评委对每个决策点逐项打分（0-100分制）。整体方案得分为各决策点评分的加权平均值，由组织者自动计算，不再由评委单独打分\n\n**评审维度示例**：方案逻辑合理性、完整性、潜在风险分析、实施资源消耗预估、可行性评审、优化建议。\n\n**评审时长**：不超过120秒/轮。若120秒仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：组织者汇总各委员对所有决策点的评分结果，标记出未通过（评分 < 阈值）的决策点，提取为「未通过决策点清单」。\n\n---\n\n### 第 2 轮：分歧讨论（Dispute Discussion）\n\n**动作说明**：组织者将第1轮中标记为「未通过」的决策点整理为「未通过决策点清单」，仅将这些未通过项下发给所有委员，要求评委**只针对这些未通过决策点**进行讨论并重新评分。已通过的决策点不再讨论。\n\n**评审时长**：不超过120秒/轮。若120秒仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：汇总本轮重新评分结果，若所有决策点全员通过则直接输出最终报告；若仍有未通过项则进入下一轮。\n\n---\n\n### 最后一轮：最终辩论 (Final Duel)\n\n**特殊说明**：无论配置多少轮，最终辩论始终是**最后一轮**。\n- 3轮配置：第3轮为最终辩论\n- 5轮配置：第5轮为最终辩论\n- N轮配置：第N轮为最终辩论\n\n**动作说明**：经过分歧讨论后，仍有决策点未通过时启用。组织者将仍未通过的决策点再次下发给所有委员，要求评委进行最后一轮讨论和重新评分。\n\n**评审时长**：不超过120秒/轮。若120秒仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：若所有决策点全员通过，输出最终报告；若仍有未通过项，将未通过项标注状态后输出最终报告。\n\n---\n\n### 轮次扩展说明\n\n| 扩展类型 | 说明 |\n|:---|:---|\n| 增加轮次（4轮及以上） | 倒数第2轮参照「分歧讨论」；最后一轮参照「最终辩论」；中间轮次循环分歧讨论 |\n| 减少轮次（3轮以下） | 第1轮参照「独立评估」；第2轮直接输出最终报告（视为最后一轮） |\n\n---\n\n## 自动化通知规则\n\n为防止评委评审时间过长导致用户对当前状态不知情，组织者必须在每轮评审开始前主动通知用户：\n\n- **进入第1轮**：通知用户「各位评委已开始独立评估，请稍候」\n- **进入第2轮**：通知用户「各位评委已开始分歧讨论，请稍候」\n- **进入最后一轮**：通知用户「各位评委已开始最终辩论，请稍候」\n- **输出最终报告**：通知用户「最终报告已生成」\n\n---\n\n## 中间结果输出规则（webchat 适配）\n\n当环境为 webchat 时，用户无法像 Feishu thread 那样看到有序的中间消息。为提升用户体验，**每轮结束后应立即向用户展示本轮汇总**，而不是等到全部轮次结束后再统一输出。\n\n| 环境 | 中间结果处理 |\n|:---|:---|\n| **webchat** | 每轮结束后立即输出本轮汇总矩阵 + 判定结果 + 提示语，然后继续下一轮 |\n| **Feishu / Discord（thread 模式）** | 按原流程执行，threaded 消息天然有序，无需特殊处理 |\n| **Unknown / 不确定** | 默认视为 webchat，执行中间输出规则 |\n\n**webchat 环境每轮输出模板**：\n```\n## 📊 第 X 轮评分汇总\n[本轮评分矩阵]\n\n## ⚖️ 第 X 轮判定结果\n[通过/未通过/进入下一轮]\n\n[未通过决策点清单（如有）]\n\n**进入第 Y 轮：[轮次名称]。各位评委已开始[环节]，请稍候。**\n```\n\n---\n\n## 通过判定规则\n\n每轮结束后**必须**进行判定，不可跳过。\n\n**【核心概念：决策点】**\n每个方案的评审内容拆解为若干「决策点」，评委对每个决策点评分（0-100分）。通过判定基于决策点级别进行。\n\n**【判定方式】（默认全票通过制）**\n- **全票通过**（默认）：所有评委在所有决策点上的评分均需≥阈值，任一评委在任一决策点低于阈值即判定为不通过\n- **均分通过**：所有评委在所有决策点上的平均分均≥阈值\n- **多数票通过**：超过半数的评委在所有决策点上的评分均≥阈值\n\n**通过判定逻辑**：\n\n| 轮次 | 判定时机 | 判定结果 | 后续动作 |\n|:---|:---|:---|:---|\n| 第1轮结束后 | 汇总所有决策点评分 | **全员通过**（所有决策点所有评委均≥阈值） | 直接输出最终报告 |\n| 第1轮结束后 | 汇总所有决策点评分 | **非全员通过**（存在未通过决策点） | 提取未通过的决策点，进入第2轮 |\n| 第2轮结束后 | 汇总未通过决策点 | **全员通过** | 直接输出最终报告 |\n| 第2轮结束后 | 汇总未通过决策点 | **仍有未通过** | 进入最后一轮最终辩论 |\n| 最后一轮结束后 | 汇总未通过决策点 | **全员通过** | 直接输出最终报告 |\n| 最后一轮结束后 | 汇总未通过决策点 | **仍有未通过** | 输出最终报告，标注各决策点状态 |\n| 任意轮超限 | — | — | 标记超时评委，以已返回结果继续汇总 |\n\n**各轮次讨论范围**：\n\n| 轮次 | 讨论范围 |\n|:---|:---|\n| 第2轮（分歧讨论） | 仅讨论上一轮中未通过的决策点 |\n| 第3轮（最终辩论） | 仅讨论第2轮中仍未通过的决策点 |\n| 后续轮次 | 循环「分歧讨论」逻辑，直至所有决策点全员通过或已达最后一轮 |\n\n---\n\n## 最终输出：6段式共识报告\n\n决策完成后，输出标准化报告，包含：\n\n1. **投票记录** — 量化评分矩阵（含决策点级别评分）\n2. **汇总说明** — 各委员核心论点（注明大模型）\n3. **执行方案** — 结论版 + 说明版\n4. **决策点通过清单** — 各决策点通过状态\n5. **结论摘要** — 一句话裁定 + 置信度\n6. **风险提示** — 主要风险与缓解建议\n\n**决策点通过状态标注**：\n\n| 状态 | 含义 | 标注 |\n|:---|:---|:---|\n| 全员通过 | 所有评委在该决策点评分均≥阈值 | ✅ 绿色 — 通过 |\n| 待决策 | 推荐顺序一致但有评委评分低于阈值 | 🟡 黄色 — 待用户确认 |\n| 有分歧 | 推荐顺序不一致 | 🔴 红色 — 有分歧，附原因 |\n\n---\n\n## 子Agent结果回收机制（重要）\n\n### Push-based 结果回收流程\n\n组织者使用 **Push-based（推送式）** 而非轮询来回收子Agent结果：\n\n```\nSpawn 子Agent → sessions_yield 挂起 → 收到 subagent_announce 事件 → 提取结果 → 更新跟踪\n```\n\n**关键步骤**：\n\n1. **Spawn 子Agent** 后，立即调用 `sessions_yield` 主动结束当前 turn\n2. **等待** OpenClaw 以 inter-session message 形式推送 `subagent_announce` 事件\n3. **识别** 消息中的 `BEGIN_UNTRUSTED_CHILD_RESULT` 块，提取评审结果\n4. **记录** 结果到会话上下文（见「轻量级状态跟踪」章节）\n5. **检查** 是否所有评委都已返回：\n   - 是 → 进入下一轮或输出最终报告\n   - 否 → 继续等待（再次 yield）\n\n⚠️ **禁止行为（铁律）**：\n\n1. **严禁跳过 sessions_yield**：spawn 评委后必须立即调用 `sessions_yield` 挂起，等待 `subagent_announce` 事件触发结果回流。严禁在结果未到达前主动调用任何工具查询结果（含 `sessions_history`）。\n\n2. **严禁跳轮宣布**：宣布进入下一轮前，必须先完成本轮所有评委结果的收集。未收集完全就宣布进入下一轮等同于「规则执行偏差」。\n\n3. **严禁用 sessions_list / subagents list / exec sleep 做状态探测**：任何形式的主动轮询均禁止，push-based 事件驱动是唯一合法结果回收方式。\n\n4. **严禁在结果未全部到达时输出最终报告**：必须等待所有评委返回，或等待超时后以已到达结果继续，同时标记未到达评委为「超时-未提交」。\n\n### 标准调用示例\n\n**Webchat 环境（推荐默认）**：\n```yaml\nSessionsSpawn:\n  runtime: \"subagent\"\n  mode: \"run\"\n  model: \"n1n/gpt-5.4\"\n  task: \"评审任务...\"\n  timeoutSeconds: 180\n```\n\n**Feishu/Discord 环境**（支持 thread 的通道）：\n```yaml\nSessionsSpawn:\n  runtime: \"acp\"\n  mode: \"session\"\n  thread: true\n  streamTo: \"parent\"\n  model: \"n1n/gpt-5.4\"\n  task: \"评审任务...\"\n  timeoutSeconds: 180\n```\n\n### 超时处理\n\n单轮限时 120 秒（2分钟）。若超时：\n- 标记该评委为「超时-未提交」\n- 以已返回的评委结果继续汇总\n- 最终报告中注明超时评委\n\n---\n\n## 轻量级状态跟踪（替代状态文件）\n\n组织者使用**会话上下文变量**跟踪评审状态，无需维护持久化状态文件。\n\n**状态跟踪格式**（每轮开始时初始化）：\n```\n[本轮状态跟踪]\nRound: {1/2/3}\nSpawned: {N}\nReceived: {M} ({评委A} ✅, {评委B} ✅)\nPending: {N-M} ({评委C} ⏳)\n```\n\n每收到一个子Agent结果，更新此跟踪块。全部完成后清空并进入下一轮。\n\n---\n\n## 异常处理规则\n\n以下3类异常的处理标准：\n\n| 异常类型 | 处理标准 |\n|:---|:---|\n| 状态同步错误 | 组织者检测到会话上下文中的状态跟踪与实际回收结果不符时，应立即中止当前流程，通知用户并提示重新发起决策 |\n| 规则执行偏差 | 组织者在执行中发现某委员的评审结果违反本skill规定的流程（如跳轮、跳过通过判定等），应要求该委员重新按规则执行，不得擅自修改委员结论 |\n| 超时处理 | 单轮超过120秒仍有评委未返回结果时，组织者对超时评委标记「超时-未提交」，以已返回的评委结果进行汇总和判定 |\n| 结果未回流 | 若 spawn 后未收到 subagent_announce 事件，检查是否使用了 sessions_yield 挂起，或 runtime/mode 配置是否正确 |\n| 框架绕过 | 组织者绕过 skill 机制私自操作（如绕 sessions_yield 直接查 history、跳轮宣布等），应立即承认并通知用户，由用户决定是重新发起还是继续 |\n\n---\n\n## 技术约束\n\n### 工具权限范围（安全白名单）\n- `Read` 工具仅限读取以下路径：\n  - skill 目录内文件（SKILL.md、references/、docs/ 等）\n  - `~/.openclaw/workspace/tmp_mmc_*.md`（临时报告文件）\n  - 严禁读取用户主目录下的隐藏文件（`.ssh/`、`.env/`、`.aws/` 等）\n- `Write` 工具仅限写入以下路径：\n  - `~/.openclaw/workspace/memory/` 目录（报告归档）\n  - `~/.openclaw/workspace/tmp_mmc_*.md`（临时报告文件）\n  - `~/Desktop/*`（最终报告输出到桌面）\n  - 禁止写入其他任何路径\n\n- 每个子会话最大执行时间：**120秒**（可通过 `timeoutSeconds` 配置，最大 300 秒）\n- 最大并发子会话数：**6个**（与委员数量上限一致）\n- ~~状态文件有效期：**24小时**~~（V1.6.0 起不再使用状态文件，改用会话上下文跟踪）\n- 报告自动归档路径：`~/.openclaw/workspace/memory/MONTHLY/mmd_<date>.md`\n\n---\n\n## 参考文档\n\n| 文档 | 说明 |\n|:---|:---|\n| [references/STATE_MACHINE.md](references/STATE_MACHINE.md) | 状态流转规则 + 异常处理（概念参考，V1.6.0 起实际使用会话上下文跟踪） |\n| [references/OUTPUT_TEMPLATE.md](references/OUTPUT_TEMPLATE.md) | 6段式报告完整模板 + 评委 Prompt |\n| [references/SCHEMA.md](references/SCHEMA.md) | 状态文件字段规范（已归档，V1.6.0 起不再强制使用） |\n| [references/TROUBLESHOOTING.md](references/TROUBLESHOOTING.md) | 常见失败模式与排查指南（V1.6.0 新增） |\n\n---\n\n## 更新日志\n\n| 版本 | 日期 | 变更内容 |\n|:---:|:---:|:---|\n| V1.7.1 | 2026-05-15 | 环境适配矩阵新增 Feishu direct chat 兜底方案说明 |\n| V1.7.0 | 2026-05-07 | 禁止行为升级为4条铁律（严禁跳过sessions_yield/严禁跳轮宣布/严禁轮询/严禁提前输出）；异常处理规则新增「框架绕过」类型 |\n| V1.6.0 | — | 改用会话上下文跟踪，废除状态文件 |\n| V1.5.0 | — | webchat环境适配 |\n\n---\n\n🏛️ **兼听则明，万模共鉴。**\n\nFile v1.7.1:docs/README.md\n\n# Multi-Model Consensus Council — Operation & User Guide\n\n## 📋 Version History\n\n| Version | Date | Changes |\n|:---|:---|:---|\n| V1.0.0 | 2026-04-19 | Initial release: core framework, 3-round convergence, 6-section report |\n| V1.1.0 | 2026-04-24 | Added English operation guide (merged with Chinese, English first); translation reviewed and approved by multi-model committee (Gemini/Doubao/GLM, avg 81/100); wording optimizations |\n| V1.1.3 | 2026-04-24 | SKILL.md and references/ converted to pure Chinese for AI readability; docs/ retains full bilingual documentation for global community |\n| V1.2.0 | 2026-04-25 | Restored public ClawHub version (was local customized); sync all files |\n| V1.2.1 | 2026-04-25 | **CRITICAL FIX**: Removed hardcoded model list (A1/A2/A4...) from SKILL.md; replaced with dynamic scan via `openclaw models list`; examples updated to use generic `[Model A/B/C]` placeholders |\n| V1.2.2 | 2026-04-25 | SYNC: Full file sync to GitHub and ClawHub after v1.2.1 hotfix |\n| V1.2.3 | 2026-04-25 | **UX FIX**: Clarify first-time user flow: auto-scan models → default to first 3 → prompt user to confirm/modify before starting decision |\n| V1.2.4 | 2026-04-25 | **CRITICAL FIX**: OUTPUT_TEMPLATE.md - remove \"匿名委员\" from all 3 round prompt templates; replaced with real-name format `{模型名称}（实名委员）` |\n| V1.2.5 | 2026-04-25 | **ENFORCEMENT**: SKILL.md - add mandatory threshold check rules; add forbidden items for skipping threshold/convergence checks; state machine now has explicit checkpoint enforcement |\n| V1.2.6 | 2026-04-25 | **Real-name Transparency**: Judges use standard model names; 3-layer architecture (organizer/judge/sub-agent); 100-point scoring; Round 0 preparation phase added |\n| V1.5.0 | 2026-04-25 | Sync all document versions to V1.5.0; renamed SCHEMA.md; full English content aligned with Chinese |\n| V1.5.1 | 2026-04-26 | Added round-start user notification rule; added convergence score threshold parameter; added exception handling rules section |\n| V1.5.2 | 2026-04-26 | Added judgment method configurable parameter; clarified unanimous-pass rule as default |\n\n---\n\n## 🎯 Plugin Overview\n\n**Multi-Model Consensus Council** is a \"Digital Think Tank\" designed for OpenClaw. It coordinates multiple large language models (e.g., GPT, Gemini, Doubao) to collaboratively analyze, debate, and synthesize decisions, effectively eliminating single-model bias.\n\n### Core Advantages\n\n- **Decentralized Decision-Making**: Cross-validates reasoning across models, ensuring rigor.\n- **Adversarial Review**: Rapidly identifies logical flaws through peer evaluation.\n- **Elastic Scaling**: Adjust the number of \"decision committee members\" anytime based on task complexity.\n- **Bias Elimination**: Multi-model independent review avoids single-AI cognitive blind spots.\n- **Quantitative Decision-Making**: 6-dimension scoring + weighted matrix, conclusions are evidence-based.\n- **Transparency & Trust**: Real-name committee system, model identity visible.\n- **Flexible Configuration**: All parameters adjustable — number of members, decision rounds, pass threshold, etc.\n\n---\n\n## ⚖️ Core Governance Principle: Identity Purity\n\nTo ensure absolute objectivity, this plugin enforces the **\"Identity Purity Principle\"**:\n\n- **Identity Transparency**: All committee members are identified by their model name (e.g., Kimi K2.6, GPT-5.4), maintaining visual transparency.\n- **No Role Assignment**: Committee members are forbidden from having preset roles such as \"architect\" or \"auditor\".\n- **No Biased Prompting**: Tasks must not include inducing prompts or identity roles targeted at committee members, preventing bias from prompt engineering.\n- **Conclusion Convergence Principle**: If a committee member calls a sub-agent based on their own judgment, they must converge the sub-agent's output within the same round before submitting their final conclusion to the organizer.\n- **No Duplicate Review Principle**: If the organizer (the current model in use) is a committee member, their submitted content counts as their Round 1 review result — no duplicate self-review is required. If the organizer is not a committee member, they are responsible only for organizing and summarizing; they do not participate in the review.\n\n---\n\n## How to Use\n\n- **Activation**: In OpenClaw, when the user's input contains 「多模型决策」 or 「多模型委员会」, the Multi-Model Consensus Council activates automatically. On first use, the system will prompt you to enter configuration mode and select the model lineup.\n- **Configuration**: Adjust parameters as needed, such as rounds, threshold, number of members, and member models.\n\n### Parameter Configuration\n\n| Command Keyword | Adjustable Parameter | Description |\n|:---|:---|:---|\n| \"换模型\" / \"Change models\" | Committee model mix | Select models from the locally connected model list |\n| \"改轮数\" / \"Change rounds\" | Decision rounds | Default 3 rounds; 2 for routine tasks; 3-6 for critical decisions |\n| \"改阈值\" / \"Change threshold\" | Judgment thresholds | **Pass threshold**: default ≥90%; **Pending threshold**: default [72%, 90%); **Reject threshold**: default <72% |\n| \"改判定方式\" / \"Change judgment method\" | Judgment method | **Unanimous-pass (default)**: all judges must meet threshold; Average-pass: average score meets threshold; Majority-pass: more than half meet threshold |\n| ~~\"收敛分差阈值\"~~ / ~~\"Convergence threshold\"~~ | ~~Score convergence threshold~~ | ~~Deprecated — replaced by decision-point pass logic~~ |\n| \"增加委员\" / \"Add member\" | Committee size | Increase number of committee members |\n| \"减少委员\" / \"Remove member\" | Committee size | Decrease number of committee members |\n| \"确认配置\" / \"Confirm config\" | Confirm current config | Confirm current parameters and begin decision |\n\n### Parameter Recommendations\n\n- **Model Selection**: Recommend at least 1 model with strong logical reasoning and 1 model with strong Chinese-language comprehension.\n- **Round Recommendations**: Routine tasks: 2 rounds; Decisions involving architecture, funding, or core rules: 3-5 rounds.\n- **Threshold Settings**:\n  - **Strict**: ≥95% (near-unanimous)\n  - **Efficient**: ≥75% (majority)\n\n---\n\n## Decision Process (3-Round Convergence Mechanism)\n\nThe Multi-Model Consensus Council adopts a 3-round convergence mechanism, with each round using a **100-point scoring system**, ultimately producing a decision result.\n\n---\n\n### Round 0: Preparation (Preparation)\n\n**Task**: Upon receiving a decision task, the **organizer** (the current model in use) summarizes the user background, requirements, and the proposal to initiate the decision application. Content format:\n\n- **Project Name**: {Project Name}\n- **Project Background**: {Project Background}\n- **Project Requirements**: {Project Requirements}\n- **Proposal Under Review**: {Proposal Under Review}\n- **Decision Members**: {Decision Members: pull from configured model lineup}\n- **Decision Rounds**: {Decision Rounds: pull from configured round count}\n- **Threshold Settings**: {Threshold Settings: pull from configured threshold}\n- **Submit Decision**: {Submit whether the user wants to begin decision}\n\n---\n\n### Round 1: Independent Evaluation (Blind Evaluation)\n\n**Action**: Upon receiving the decision confirmation, the organizer distributes the preparation content to each selected model's independent sub-session. Each committee member conducts closed-door review **without any awareness of other members' opinions**.\n\n**Review Dimensions**: Logical completeness, risk analysis, resource cost estimation, feasibility review, optimization suggestions — all scored on a **100-point scale**.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: The organizer summarizes each committee member's initial review report, voting results, scores, and disagreement points or controversy areas.\n\n---\n\n### Convergence Rounds (All rounds except the final round)\n\n**Action**: The organizer summarizes and finalizes all results agreed upon by the full committee — no further discussion on those points. Extracts all unresolved points from the previous round (items below threshold), organizes them into a \"dispute list,\" redistributing to all committee members for re-evaluation and re-stance **only on these unresolved items**, ultimately producing this round's decision results.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: Disputes largely converge, producing this round's decision results. **If all committee members approve the decision (no new dispute points), skip the next round and output the final report.** If disputes remain, proceed to the next convergence round.\n\n---\n\n### Final Round: Final Duel\n\n**Special Note**: Regardless of how many rounds are configured, the Final Duel is **always the last round**.\n- 3-round config: Round 3 is the Final Duel\n- 5-round config: Round 5 is the Final Duel\n- N-round config: Round N is the Final Duel\n\n**Action**: After multiple convergence rounds, if deep conflicts remain unresolved, this round is activated. The organizer redistributes these \"deep conflict points\" to all committee members, simplified into binary options — \"Yes/No\" or \"Option A/Option B\" — and requires all committee members to make their final logical stance.\n\n**Pass Condition**: **Consensus Selection** — all judges have identical recommendation order = pass; threshold is NOT a hard requirement.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: Produces the final decision results. Generates the final quantitative voting matrix and automatically produces a 6-section standard decision report.\n\n---\n\n### Round Extension (for non-default round settings)\n\n| Extension Type | Description |\n|:---|:---|\n| Increase rounds (4 or more) | Round 4 follows \"Consensus Sync\"; Round 5 follows \"Final Duel\"; and so on |\n| Decrease rounds (2 rounds) | Round 1 follows \"Independent Evaluation\"; Round 2 directly outputs the final report |\n\n---\n\n## Final Output: 6-Section Consensus Report\n\nThe decision output is a standardized report containing:\n\n1. **Voting Record** — Quantitative scoring matrix\n2. **Summary** — Each committee member's core arguments (with model name noted)\n3. **Execution Plan** — Conclusion version + detailed version\n4. **Unresolved List** — Pending items and responsible parties\n5. **Conclusion Summary** — One-sentence verdict + confidence level\n6. **Risk Advisory** — Key risks and mitigation recommendations\n\n---\n\n## Judgment Method\n\nThree configurable judgment methods (default: unanimous-pass):\n\n| Method | Rule |\n|:---|:---|\n| **Unanimous-pass (default)** | All judges must individually score ≥ threshold |\n| **Average-pass** | Average score of all judges ≥ threshold |\n| **Majority-pass** | More than half of judges score ≥ threshold |\n\n---\n\n## Exception Handling Rules\n\n- **Timeout**: If sub-sessions are not all recovered within 3 minutes, the organizer terminates and returns available results.\n- **Model Failure**: If a judge model fails to respond, the organizer marks it as \"未响应\" and proceeds with available results.\n- **Score Spread**: If max score spread exceeds the convergence threshold, automatically trigger the next round.\n\n---\n*Version: V1.5.2*\n*Developer: Zeekr0808*\n*Email: Zeekr0808@outlook.com*\n\n---\n\n# 多模型决策委员会 — 操作与使用指南\n\n## 版本变更记录\n\n| 版本 | 日期 | 更新内容 |\n|:---|:---|:---|\n| V1.0.0 | 2026-04-19 | 初始版本，包含核心框架、3轮收敛机制、6段式报告 |\n| V1.1.0 | 2026-04-24 | **新增英文操作文档**（与中文合并，英文在上）；经多模型委员会翻译质量审核通过（均分81/100） |\n| V1.1.3 | 2026-04-24 | SKILL.md和references/转为纯中文（精简）；docs/保留双语完整文档（面向全球用户） |\n| V1.2.0 | 2026-04-25 | 恢复为公共ClawHub版本（曾为本地定制版）；同步所有文件 |\n| V1.2.1 | 2026-04-25 | **关键修复**：删除SKILL.md中硬编码的模型列表（A1/A2/A4...）；改为执行时动态扫描通过`openclaw models list`获取实时模型列表；示例中的模型引用改为通用占位符`[模型A/B/C]` |\n| V1.2.2 | 2026-04-25 | 同步：v1.2.1热修复后全文件同步到GitHub和ClawHub |\n| V1.2.3 | 2026-04-25 | **用户体验修复**：明确首次使用流程：自动扫描模型 → 默认前3个 → 提示用户确认或修改评委后再开始决策 |\n| V1.2.4 | 2026-04-25 | **关键修复**：OUTPUT_TEMPLATE.md - 删除所有3轮prompt模板中的「匿名委员」，改为「{模型名称}（实名委员）」格式 |\n| V1.2.5 | 2026-04-25 | **强制判定机制**：SKILL.md - 新增「强制判定规则」章节，明确 Round 1/2 后的阈值判定和收敛判定为不可跳过步骤；禁止行为新增跳项判定禁止规则 |\n| V1.2.6 | 2026-04-25 | **实名显性化**：评委以标准大模型名称显名，透明可溯源；三层架构（组织者/评委/子Agent）；100分制评分；第0轮准备阶段 |\n| V1.5.0 | 2026-04-25 | 同步所有文档版本号至V1.5.0；英文内容全量对齐中文 |\n| V1.5.1 | 2026-04-26 | 新增「每轮开始前通知用户」规则；新增「收敛分差阈值」可配置参数；新增「异常处理规则」章节 |\n| V1.5.2 | 2026-04-26 | 新增「判定方式」可配置参数，明确全票通过制为默认判定规则 |\n| **V1.6.2** | **2026-04-27** | **核心逻辑重构**：取消「收敛」概念，引入「决策点」级别评审；通过判定改为全员通过/非全员通过；「收敛讨论」改为「分歧讨论」；最终报告新增决策点通过状态标注（✅绿色/🟡黄色/🔴红色） |\n| **V1.6.3** | **2026-04-28** | **触发词精确化**：触发条件由长句改为精确词组「多模型决策」「多模型委员会」；**webchat 中间输出适配**：每轮结束后立即向用户展示本轮汇总，不再等待全部轮次结束后统一输出；同步更新所有文档 |\n| **V1.6.5** | **2026-04-30** | **安全加固**：移除 allowed-tools 中未使用的 Exec；最大并发子Agent从13降至6；新增 Write 工具路径限制（仅限 memory 目录）；TROUBLESHOOTING.md 新增模型路由偏差提前终止规则 |\n| **V1.6.6** | **2026-04-30** | **安全修复**：将`openclaw.json`文件扫描改为`openclaw models list`命令，避免暴露 provider token；修正委员会规模描述从3-13为3-6个模型；SKILL.md 新增 Read/Write 工具路径白名单声明 |\n| **V1.6.7** | **2026-04-30** | **可审计性修复**：明确工具使用分工——组织者使用 Read/Write，评委通过 task 参数接收内容且不得调用任何工具；删除\"仅供内部\"标记；消除 OUTPUT_TEMPLATE.md 中评委 spawn 的歧义表述 |\n\n---\n\n## 插件定位\n\n**多模型决策委员会** 是专为 OpenClaw 设计的\"数字智库\"。它支持调动多个不同架构的大模型（如 GPT、Gemini、豆包等）同时对同一任务进行协同思考、对抗辩论与共识合成，有效消除单模型偏见，综合性给出最优建议。\n\n## 核心优势\n\n- **去中心化决策**：交叉验证不同模型逻辑，确保方案的严谨性。\n- **对抗性评审**：通过模型间的互评，快速锁定潜在的逻辑漏洞。\n- **弹性扩容**：根据任务难度，随时增加或减少\"决策委员\"的数量。\n- **消除偏见**：多模型独立评审，避免单一AI的认知盲区。\n- **量化决策**：6维度评分 + 加权矩阵，结论有据可查。\n- **透明可信**：实名委员制、模型身份透明。\n- **灵活配置**：参数可调，如决策成员人数、决策轮次、通过阈值等，均可可自定义。\n\n---\n\n## 核心治理原则：身份纯净\n\n为了确保结果的绝对客观，本插件强制执行 **\"身份纯净原则\"**：\n- **角色透明化**：决策委员会成员均会注明其身份（大模型名称），保持可视化透明。\n- **严禁设定角色**：禁止给委员设定诸如\"架构师\"、\"审计员\"等身份标签。\n- **严禁引导提示**：任务下发时不得针对评审委员设置诱导性提示词或诱导性身份角色，防止产生的偏见。\n- **结论统一原则**：若决策委员会成员根据判断自行调用了子Agent，则必须在本轮结束前汇总子Agent的结果，形成统一结论，再向组织者输出最终结果。\n- **评审不重复原则**：若组织者（即当前使用模型）是评审委员会成员，则组织者提交内容即视同为其在本轮的评审结果，无需重复自评。若组织者不是评审委员会成员，则仅负责组织实施和汇总评委意见，不参与评审。\n\n---\n\n## 使用方法\n-  **启动**：在OpenClaw中，当用户输入的内容包含「多模型决策」或「多模型委员会」时，自动激活多模型决策委员会。首次使用时会提醒用户进入配置模式，并选择模型组合。\n- **配置**：根据需要，可配置参数，如：轮数、阈值、委员数量、委员模型等。\n\n### 参数配置\n配置参数可采用指令方式，如：\"修改配置\"、\"调整参数\"、\"增加委员\"、\"确认配置\"。\n| 指令关键词 | 可调参数 | 说明 |\n|:---|:---|:---|\n| \"换模型\" | 委员模型组合 | 从用户本地接入的模型列表中选择模型组合加入 |\n| \"改轮数\" | 决策轮数 | 默认3轮，日常事务可设为2轮，重大决策可设3-6轮 |\n| \"改阈值\" | 判定阈值 | **通过阈值**：默认≥90%；**待决策阈值**：默认[72%, 90%)；**否定阈值**：默认<72% |\n| \"改判定方式\" | 判定方式 | **全票通过（默认）**：所有评委均需≥阈值；均分通过：评委均分≥阈值；多数票通过：超过半数评委≥阈值 |\n| ~~\"收敛分差阈值\"~~ | ~~收敛分差阈值~~ | ~~已废弃~~ |\n| \"增加委员\" / \"减少委员\" | 委员数量 | 支持2-6个模型同时评审 |\n| \"确认配置\" | 确认当前配置 | 确认当前参数后开始决策 |\n\n## 参数建议\n- **模型选择**：建议至少包含 1 个具备强逻辑推理能力的模型和 1 个具备强中文语境理解能力的模型。\n- **轮次建议**：日常事务建议 2 轮；涉及架构、资金、核心规则的决策，建议 3-5 轮。\n- **阈值设定**：\n  - **严谨型**：≥95%（全票通过）\n  - **效率型**：≥75%（多数通过）\n\n---\n\n## 决策流程（3轮决策机制）\n多模型决策委员会采用 3 轮决策机制，基于决策点级别评审，每轮进行100分制评分，最终给出决策结果。具体流程如下：\n\n---\n\n### 第 0 轮：准备与框架确认 (Preparation & Framework Confirmation)\n**决策准备**：组织者（当前使用模型）接收到决策任务后，将用户背景、需求、待审方案汇总，判断评审方案类型，发起决策申请。内容格式为：\n- **项目名称**：{项目名称}\n- **项目背景**：{项目背景}\n- **项目需求**：{项目需求}\n- **待审方案**：{待审方案}\n- **评审方案类型**：{单方案评审 / 二选一评审 / 决策点拆分评审}\n- **决策点拆分及权重**（仅决策点拆分评审时填写）：\n  - 决策点1（{维度名称}）：{权重}%\n  - 决策点2（{维度名称}）：{权重}%\n  - ……\n- **决策成员**：{决策成员：调取已配置的模型组合}\n- **决策轮次**：{决策轮次：调取已配置的轮数}\n- **阈值设置**：{阈值设置：调取已配置的阈值}\n- **提交决策**：{用户确认评审类型及配置后开始决策}\n\n**评审方案类型**（由组织者根据方案复杂程度判断）：\n- **单方案评审**：简单/日常方案，评委直接对整体方案打分（0-100分制）\n- **二选一评审**：提供方案A和方案B两套完整方案，评委先选择其中一个方案（落选方案直接PASS），再对选中方案的每个决策点逐项打分（0-100分制），加权平均计算总分\n- **决策点拆分评审**：复杂/多维度方案，组织者将方案拆解为若干「决策点」，每个决策点对应一个具体评审维度。评委对每个决策点逐项打分（0-100分制）。整体方案得分为各决策点评分的加权平均值，由组织者自动计算，不再由评委单独打分\n\n---\n\n### 第 1 轮：独立评估 (Blind Evaluation)\n\n**动作说明**：组织者接收到决策确认后，将准备流程相关内容分发至各选定模型的独立子会话中。每位委员在**完全无法感知他方意见**的情况下进行背对背评审。\n\n**评审维度**：方案逻辑合理性、完整性、潜在风险分析、实施资源消耗预估、可行性评审、优化建议，并按照100分值进行评分。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：组织者汇总各委员的初评报告及投票结果及打分结果，以及分歧意见或争议点。\n\n---\n\n### 第 2 轮：分歧讨论 (Dispute Discussion)\n\n**动作说明**：组织者将第1轮中标记为「未通过」的决策点整理为「未通过决策点清单」，仅将这些未通过项下发给所有委员，要求评委**只针对这些未通过决策点**进行讨论并重新评分。已通过的决策点不再讨论。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：汇总本轮重新评分结果，若所有决策点全员通过则直接输出最终报告；若仍有未通过项则进入下一轮。\n\n---\n\n### 最后一轮：最终辩论 (Final Duel)\n\n**特殊说明**：无论配置多少轮，最终辩论始终是**最后一轮**。\n- 3轮配置：第3轮为最终辩论\n- 5轮配置：第5轮为最终辩论\n- N轮配置：第N轮为最终辩论\n\n**动作说明**：经过分歧讨论后，仍有决策点未通过时启用。组织者将仍未通过的决策点再次下发给所有委员，要求评委进行最后一轮讨论和重新评分。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：若所有决策点全员通过，输出最终报告；若仍有未通过项，将未通过项标注状态后输出最终报告。\n\n---\n\n### 轮次扩展说明\n\n| 扩展类型 | 说明 |\n|:---|:---|\n| 增加轮次（4轮及以上） | 倒数第2轮参照「分歧讨论」；最后一轮参照「最终辩论」；中间轮次循环分歧讨论 |\n| 减少轮次（2轮） | 第1轮参照\"独立评估\"，第2轮直接输出最终报告 |\n\n---\n\n## 最终输出：6段式共识报告\n\n决策完成后，输出标准化报告，包含：\n\n1. **投票记录** — 量化评分矩阵（含决策点级别评分）\n2. **汇总说明** — 各委员核心论点（注明大模型）\n3. **执行方案** — 结论版 + 说明版\n4. **决策点通过清单** — 各决策点通过状态\n5. **结论摘要** — 一句话裁定 + 置信度\n6. **风险提示** — 主要风险与缓解建议\n\n**决策点通过状态标注**：\n\n| 状态 | 含义 | 标注 |\n|:---|:---|:---|\n| 全员通过 | 所有评委在该决策点评分均≥阈值 | ✅ 绿色 — 通过 |\n| 待决策 | 推荐顺序一致但有评委评分低于阈值 | 🟡 黄色 — 待用户确认 |\n| 有分歧 | 推荐顺序不一致 | 🔴 红色 — 有分歧，附原因 |\n\n---\n\n## 判定方式\n\n三种可配置的判定方式（默认：全票通过）：\n\n| 判定方式 | 规则 |\n|:---|:---|\n| **全票通过（默认）** | 所有评委均需≥阈值 |\n| **均分通过** | 评委均分≥阈值 |\n| **多数票通过** | 超过半数评委≥阈值 |\n\n---\n\n## 异常处理规则\n\n- **超时**：3分钟内未回收所有子会话，组织者终止并返回已收集结果\n- **模型故障**：评委模型无响应时，标记为\"未响应\"，以已收集结果继续进行\n- **超时处理**：3分钟内未回收所有子会话，标记超时评委，以已返回结果继续汇总\n\n---\n*版本: V1.6.7*\n*开发者: Zeekr0808*\n*邮箱: Zeekr0808@outlook.com*\n\nFile v1.7.1:README.md\n\n# Multi-Model Consensus Council — Operation & User Guide\n\n## 📋 Version History\n\n| Version | Date | Changes |\n|:---|:---|:---|\n| V1.0.0 | 2026-04-19 | Initial release: core framework, 3-round convergence, 6-section report |\n| V1.1.0 | 2026-04-24 | Added English operation guide (merged with Chinese, English first); translation reviewed and approved by multi-model committee (Gemini/Doubao/GLM, avg 81/100); wording optimizations |\n| V1.1.3 | 2026-04-24 | SKILL.md and references/ converted to pure Chinese for AI readability; docs/ retains full bilingual documentation for global community |\n| V1.2.0 | 2026-04-25 | Restored public ClawHub version (was local customized); sync all files |\n| V1.2.1 | 2026-04-25 | **CRITICAL FIX**: Removed hardcoded model list (A1/A2/A4...) from SKILL.md; replaced with dynamic scan via `openclaw models list`; examples updated to use generic `[Model A/B/C]` placeholders |\n| V1.2.2 | 2026-04-25 | SYNC: Full file sync to GitHub and ClawHub after v1.2.1 hotfix |\n| V1.2.3 | 2026-04-25 | **UX FIX**: Clarify first-time user flow: auto-scan models → default to first 3 → prompt user to confirm/modify before starting decision |\n| V1.2.4 | 2026-04-25 | **CRITICAL FIX**: OUTPUT_TEMPLATE.md - remove \"匿名委员\" from all 3 round prompt templates; replaced with real-name format `{模型名称}（实名委员）` |\n| V1.2.5 | 2026-04-25 | **ENFORCEMENT**: SKILL.md - add mandatory threshold check rules; add forbidden items for skipping threshold/convergence checks; state machine now has explicit checkpoint enforcement |\n| V1.2.6 | 2026-04-25 | **Real-name Transparency**: Judges use standard model names; 3-layer architecture (organizer/judge/sub-agent); 100-point scoring; Round 0 preparation phase added |\n| V1.5.0 | 2026-04-25 | Sync all document versions to V1.5.0; renamed SCHEMA.md; full English content aligned with Chinese |\n| V1.5.1 | 2026-04-26 | Added round-start user notification rule; added convergence score threshold parameter; added exception handling rules section |\n| V1.5.2 | 2026-04-26 | Added judgment method configurable parameter; clarified unanimous-pass rule as default |\n| **V1.6.0** | **2026-04-27** | **MAJOR FIX**: Rewrote sub-agent result recovery; added runtime environment adaptation matrix; Push-based waiting flow; runtime self-check; simplified state management; added TROUBLESHOOTING.md |\n| **V1.6.1** | **2026-04-27** | **Clarified judgment rules**: Separated \"Pass\" vs \"Convergence\" vs \"Consensus Selection\"; clarified convergence discussion scope; renamed \"Round 3\" to \"Final Round\"; added Consensus Selection for final round |\n| **V1.6.2** | **2026-04-27** | **Core logic refactored**: Removed \"Convergence\"; introduced \"Decision Point\" level review; \"Convergence Discussion\" renamed to \"Dispute Discussion\"; final report includes decision point status (✅/🟡/🔴) |\n| **V1.6.5** | **2026-04-30** | **Security Hardening**: Removed unused `Exec` from allowed-tools; reduced max concurrent sub-agents from 13 to 6; added `Write` tool path restriction (memory directory only); added model-routing early termination rule in TROUBLESHOOTING.md |\n| **V1.6.6** | **2026-04-30** | **Security Fix**: Replaced `openclaw.json` file scan with `openclaw models list` command to avoid exposing provider tokens; corrected committee size description from 3-13 to 3-6 models; added Read/Write tool path whitelist in SKILL.md |\n| **V1.6.7** | **2026-04-30** | **Auditability Fix**: Clarified tool usage scope—organizers use Read/Write, judges receive content via task parameter and must not call any tools; removed contradictory \"internal-only\" markings; eliminated ambiguous judge spawn instructions in OUTPUT_TEMPLATE.md |\n\n---\n\n## 🎯 Plugin Overview\n\n**Multi-Model Consensus Council** is a \"Digital Think Tank\" designed for OpenClaw. It coordinates multiple large language models (e.g., GPT, Gemini, Doubao) to collaboratively analyze, debate, and synthesize decisions, effectively eliminating single-model bias.\n\n### Core Advantages\n\n- **Decentralized Decision-Making**: Cross-validates reasoning across models, ensuring rigor.\n- **Adversarial Review**: Rapidly identifies logical flaws through peer evaluation.\n- **Elastic Scaling**: Adjust the number of \"decision committee members\" anytime based on task complexity.\n- **Bias Elimination**: Multi-model independent review avoids single-AI cognitive blind spots.\n- **Quantitative Decision-Making**: 6-dimension scoring + weighted matrix, conclusions are evidence-based.\n- **Transparency & Trust**: Real-name committee system, model identity visible.\n- **Flexible Configuration**: All parameters adjustable — number of members, decision rounds, pass threshold, etc.\n\n---\n\n## ⚖️ Core Governance Principle: Identity Purity\n\nTo ensure absolute objectivity, this plugin enforces the **\"Identity Purity Principle\"**:\n\n- **Identity Transparency**: All committee members are identified by their model name (e.g., Kimi K2.6, GPT-5.4), maintaining visual transparency.\n- **No Role Assignment**: Committee members are forbidden from having preset roles such as \"architect\" or \"auditor\".\n- **No Biased Prompting**: Tasks must not include inducing prompts or identity roles targeted at committee members, preventing bias from prompt engineering.\n- **No Sub-Agent Delegation**: Judges are strictly prohibited from spawning sub-agents. They must think independently and output their own conclusions. Complex tasks are decomposed and distributed by the organizer.\n- **No Duplicate Review Principle**: If the organizer (the current model in use) is a committee member, their submitted content counts as their Round 1 review result — no duplicate self-review is required. If the organizer is not a committee member, they are responsible only for organizing and summarizing; they do not participate in the review.\n\n---\n\n## How to Use\n\n- **Activation**: In OpenClaw, type \"启动多模型决策委员会\" (Start Multi-Model Consensus Council), \"审议这个方案\" (Review this plan), \"对这个方案进行决策\" (Make a decision on this plan), or \"启动决策\" (Start decision) to activate. On first use, the system will prompt you to enter configuration mode and select the model lineup.\n- **Configuration**: Adjust parameters as needed, such as rounds, threshold, number of members, and member models.\n\n### Parameter Configuration\n\n| Command Keyword | Adjustable Parameter | Description |\n|:---|:---|:---|\n| \"换模型\" / \"Change models\" | Committee model mix | Select models from the locally connected model list |\n| \"改轮数\" / \"Change rounds\" | Decision rounds | Default 3 rounds; 2 for routine tasks; 3-6 for critical decisions |\n| \"改阈值\" / \"Change threshold\" | Judgment thresholds | **Pass threshold**: default ≥90%; **Pending threshold**: default [72%, 90%); **Reject threshold**: default <72% |\n| \"增加委员\" / \"Add member\" | Committee size | Increase number of committee members |\n| \"减少委员\" / \"Remove member\" | Committee size | Decrease number of committee members |\n| \"确认配置\" / \"Confirm config\" | Confirm current config | Confirm current parameters and begin decision |\n\n### Parameter Recommendations\n\n- **Model Selection**: Recommend at least 1 model with strong logical reasoning and 1 model with strong Chinese-language comprehension.\n- **Round Recommendations**: Routine tasks: 2 rounds; Decisions involving architecture, funding, or core rules: 3-5 rounds.\n- **Threshold Settings**:\n  - **Strict**: ≥95% (near-unanimous)\n  - **Efficient**: ≥75% (majority)\n\n---\n\n## Decision Process (3-Round Convergence Mechanism)\n\nThe Multi-Model Consensus Council adopts a 3-round convergence mechanism, with each round using a **100-point scoring system**, ultimately producing a decision result.\n\n---\n\n### Round 0: Preparation (Preparation)\n\n**Task**: Upon receiving a decision task, the **organizer** (the current model in use) summarizes the user background, requirements, and the proposal to initiate the decision application. Content format:\n\n- **Project Name**: {Project Name}\n- **Project Background**: {Project Background}\n- **Project Requirements**: {Project Requirements}\n- **Proposal Under Review**: {Proposal Under Review}\n- **Decision Members**: {Decision Members: pull from configured model lineup}\n- **Decision Rounds**: {Decision Rounds: pull from configured round count}\n- **Threshold Settings**: {Threshold Settings: pull from configured threshold}\n- **Submit Decision**: {Submit whether the user wants to begin decision}\n\n---\n\n### Round 1: Independent Evaluation (Blind Evaluation)\n\n**Action**: Upon receiving the decision confirmation, the organizer distributes the preparation content to each selected model's independent sub-session. Each committee member conducts closed-door review **without any awareness of other members' opinions**.\n\n**Review Dimensions**: Logical completeness, risk analysis, resource cost estimation, feasibility review, optimization suggestions — all scored on a **100-point scale**.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: The organizer summarizes each committee member's initial review report, voting results, scores, and disagreement points or controversy areas.\n\n---\n\n### Round 2: Dispute Discussion\n\n**Action**: The organizer extracts all decision points that did NOT achieve full pass in Round 1 (any judge scoring below threshold) and compiles them into an \"unresolved decision points list.\" These unresolved points only are redistributed to all committee members for re-discussion and re-scoring. Points already with full pass are frozen and not re-discussed.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: Summarizes re-scoring results. If all previously unresolved points now achieve full pass, output the final report. If any remain unresolved, proceed to the final round.\n\n---\n\n### Final Round: Final Duel\n\n**Special Note**: Regardless of how many rounds are configured, the Final Duel is **always the last round**.\n- 3-round config: Round 3 is the Final Duel\n- 5-round config: Round 5 is the Final Duel\n- N-round config: Round N is the Final Duel\n\n**Action**: After dispute discussion, if any decision points still have judges scoring below threshold, those unresolved points are redistributed to all committee members for a final round of discussion and re-scoring.\n\n**Review Duration**: No more than 3 minutes per round. If not all sub-sessions are recovered after 3 minutes, the organizer terminates the process and returns results.\n\n**Core Output**: If all decision points now achieve full pass, output the final report. If any remain unresolved, output the final report with each decision point's status tagged: ✅ Full Pass / 🟡 Pending User Decision / 🔴 Disputed.\n\n---\n\n### Round Extension (for non-default round settings)\n\n| Extension Type | Description |\n|:---|:---|\n| Increase rounds (4 or more) | 2nd-to-last round follows \"Dispute Discussion\"; last round follows \"Final Duel\"; middle rounds loop dispute discussion |\n| Decrease rounds (2 rounds) | Round 1 follows \"Independent Evaluation\"; Round 2 directly outputs the final report (treated as final round) |\n\n---\n\n## Final Output: 6-Section Consensus Report\n\nThe decision output is a standardized report containing:\n\n1. **Voting Record** — Quantitative scoring matrix (by decision point)\n2. **Summary** — Each committee member's core arguments (with model name noted)\n3. **Execution Plan** — Conclusion version + detailed version\n4. **Decision Point Status** — Each decision point's pass status\n5. **Conclusion Summary** — One-sentence verdict + confidence level\n6. **Risk Advisory** — Key risks and mitigation recommendations\n\n**Decision Point Status Tags**:\n| Status | Meaning | Tag |\n|:---|:---|:---|\n| Full Pass | All judges scored ≥ threshold on this point | ✅ Green — Passed |\n| Pending Decision | All judges chose same option but some scores below threshold | 🟡 Yellow — Pending User Confirmation |\n| Disputed | Judges chose different options | 🔴 Red — Disputed, reason noted |\n\n---\n*Version: V1.6.7*\n*Developer: Zeekr0808*\n*Email: Zeekr0808@outlook.com*\n\n---\n\n# 多模型决策委员会 — 操作与使用指南\n\n## 版本变更记录\n\n| 版本 | 日期 | 更新内容 |\n|:---|:---|:---|\n| V1.0.0 | 2026-04-19 | 初始版本，包含核心框架、3轮收敛机制、6段式报告 |\n| V1.1.0 | 2026-04-24 | **新增英文操作文档**（与中文合并，英文在上）；经多模型委员会翻译质量审核通过（均分81/100） |\n| V1.1.3 | 2026-04-24 | SKILL.md和references/转为纯中文（精简）；docs/保留双语完整文档（面向全球用户） |\n| V1.2.0 | 2026-04-25 | 恢复为公共ClawHub版本（曾为本地定制版）；同步所有文件 |\n| V1.2.1 | 2026-04-25 | **关键修复**：删除SKILL.md中硬编码的模型列表（A1/A2/A4...）；改为执行时动态扫描通过`openclaw models list`获取实时模型列表；示例中的模型引用改为通用占位符`[模型A/B/C]` |\n| V1.2.2 | 2026-04-25 | 同步：v1.2.1热修复后全文件同步到GitHub和ClawHub |\n| V1.2.3 | 2026-04-25 | **用户体验修复**：明确首次使用流程：自动扫描模型 → 默认前3个 → 提示用户确认或修改评委后再开始决策 |\n| V1.2.4 | 2026-04-25 | **关键修复**：OUTPUT_TEMPLATE.md - 删除所有3轮prompt模板中的「匿名委员」，改为「{模型名称}（实名委员）」格式 |\n| V1.2.5 | 2026-04-25 | **强制判定机制**：SKILL.md - 新增「强制判定规则」章节，明确 Round 1/2 后的阈值判定和收敛判定为不可跳过步骤；禁止行为新增跳项判定禁止规则 |\n| V1.2.6 | 2026-04-25 | **实名显性化**：评委以标准大模型名称显名，透明可溯源；三层架构（组织者/评委/子Agent）；100分制评分；第0轮准备阶段 |\n| V1.5.0 | 2026-04-25 | 同步所有文档版本号至V1.5.0；英文内容全量对齐中文 |\n| V1.5.1 | 2026-04-26 | 新增「每轮开始前通知用户」规则；新增「收敛分差阈值」可配置参数；新增「异常处理规则」章节 |\n| V1.5.2 | 2026-04-26 | 新增「判定方式」可配置参数，明确全票通过制为默认判定规则 |\n| **V1.6.0** | **2026-04-27** | **重大修复**：重写子Agent结果回收机制，新增运行时环境适配矩阵；新增Push-based等待流程；新增运行时自检；简化状态管理；新增TROUBLESHOOTING.md |\n| **V1.6.1** | **2026-04-27** | **判定规则澄清**：区分「通过」「收敛」「一致性选择」三个独立概念；澄清收敛讨论范围；「第3轮」更名为「最后一轮」；新增最终轮「一致性选择」机制 |\n| **V1.6.2** | **2026-04-27** | **核心逻辑重构**：取消「收敛」概念，引入「决策点」级别评审；通过判定改为全员通过/非全员通过；「收敛讨论」改为「分歧讨论」；最终报告新增决策点通过状态标注（✅绿色/🟡黄色/🔴红色） |\n| **V1.6.5** | **2026-04-30** | **安全加固**：移除 allowed-tools 中未使用的 Exec；最大并发子Agent从13降至6；新增 Write 工具路径限制（仅限 memory 目录）；TROUBLESHOOTING.md 新增模型路由偏差提前终止规则 |\n| **V1.6.6** | **2026-04-30** | **安全修复**：将`openclaw.json`文件扫描改为`openclaw models list`命令，避免暴露 provider token；修正委员会规模描述从3-13为3-6个模型；SKILL.md 新增 Read/Write 工具路径白名单声明 |\n| **V1.6.7** | **2026-04-30** | **可审计性修复**：明确工具使用分工——组织者使用 Read/Write，评委通过 task 参数接收内容且不得调用任何工具；删除\"仅供内部\"标记；消除 OUTPUT_TEMPLATE.md 中评委 spawn 的歧义表述 |\n\n---\n\n## 插件定位\n\n**多模型决策委员会** 是专为 OpenClaw 设计的\"数字智库\"。它支持调动多个不同架构的大模型（如 GPT、Gemini、豆包等）同时对同一任务进行协同思考、对抗辩论与共识合成，有效消除单模型偏见，综合性给出最优建议。\n\n## 核心优势\n\n- **去中心化决策**：交叉验证不同模型逻辑，确保方案的严谨性。\n- **对抗性评审**：通过模型间的互评，快速锁定潜在的逻辑漏洞。\n- **弹性扩容**：根据任务难度，随时增加或减少\"决策委员\"的数量。\n- **消除偏见**：多模型独立评审，避免单一AI的认知盲区。\n- **量化决策**：6维度评分 + 加权矩阵，结论有据可查。\n- **透明可信**：实名委员制、模型身份透明。\n- **灵活配置**：参数可调，如决策成员人数、决策轮次、通过阈值等，均可可自定义。\n\n---\n\n## 核心治理原则：身份纯净\n\n为了确保结果的绝对客观，本插件强制执行 **\"身份纯净原则\"**：\n- **角色透明化**：决策委员会成员均会注明其身份（大模型名称），保持可视化透明。\n- **严禁设定角色**：禁止给委员设定诸如\"架构师\"、\"审计员\"等身份标签。\n- **严禁引导提示**：任务下发时不得针对评审委员设置诱导性提示词或诱导性身份角色，防止产生的偏见。\n- **严禁子Agent委托**：评委严禁spawn任何子Agent，只能独立思考和输出结论；禁止二次委托，所有子任务由组织者统一分发和回收，确保链路完全可控。\n- **评审不重复原则**：若组织者（即当前使用模型）是评审委员会成员，则组织者提交内容即视同为其在本轮的评审结果，无需重复自评。若组织者不是评审委员会成员，则仅负责组织实施和汇总评委意见，不参与评审。\n\n---\n\n## 使用方法\n-  **启动**：在OpenClaw中，输入指令\"启动多模型决策委员会\"、\"审议这个方案\"、\"对这个方案进行决策\"、\"启动决策\"即可启动多模型决策委员会。首次使用时会提醒用户进入配置模式，并选择模型组合。\n- **配置**：根据需要，可配置参数，如：轮数、阈值、委员数量、委员模型等。\n\n### 参数配置\n配置参数可采用指令方式，如：\"修改配置\"、\"调整参数\"、\"增加委员\"、\"确认配置\"。\n| 指令关键词 | 可调参数 | 说明 |\n|:---|:---|:---|\n| \"换模型\" | 委员模型组合 | 从用户本地接入的模型列表中选择模型组合加入 |\n| \"改轮数\" | 决策轮数 | 默认3轮，日常事务可设为2轮，重大决策可设3-6轮 |\n| \"改阈值\" | 判定阈值 | **通过阈值**：默认≥90%；**待决策阈值**：默认[72%, 90%)；**否定阈值**：默认<72% |\n| \"增加委员\" / \"减少委员\" | 委员数量 | 支持2-6个模型同时评审 |\n| \"确认配置\" | 确认当前配置 | 确认当前参数后开始决策 |\n\n## 参数建议\n- **模型选择**：建议至少包含 1 个具备强逻辑推理能力的模型和 1 个具备强中文语境理解能力的模型。\n- **轮次建议**：日常事务建议 2 轮；涉及架构、资金、核心规则的决策，建议 3-5 轮。\n- **阈值设定**：\n  - **严谨型**：≥95%（全票通过）\n  - **效率型**：≥75%（多数通过）\n\n---\n\n## 决策流程（3轮决策机制）\n多模型决策委员会采用 3 轮决策机制，基于决策点级别评审，每轮进行100分制评分，最终给出决策结果。具体流程如下：\n\n---\n\n### 第 0 轮：准备与框架确认 (Preparation & Framework Confirmation)\n**决策准备**：组织者（当前使用模型）接收到决策任务后，将用户背景、需求、待审方案汇总，判断评审方案类型，发起决策申请。内容格式为：\n- **项目名称**：{项目名称}\n- **项目背景**：{项目背景}\n- **项目需求**：{项目需求}\n- **待审方案**：{待审方案}\n- **评审方案类型**：{单方案评审 / 二选一评审 / 决策点拆分评审}\n- **决策点拆分及权重**（仅决策点拆分评审时填写）：\n  - 决策点1（{维度名称}）：{权重}%\n  - 决策点2（{维度名称}）：{权重}%\n  - ……\n- **决策成员**：{决策成员：调取已配置的模型组合}\n- **决策轮次**：{决策轮次：调取已配置的轮数}\n- **阈值设置**：{阈值设置：调取已配置的阈值}\n- **提交决策**：{用户确认评审类型及配置后开始决策}\n\n**评审方案类型**（由组织者根据方案复杂程度判断）：\n- **单方案评审**：简单/日常方案，评委直接对整体方案打分（0-100分制）\n- **二选一评审**：提供方案A和方案B两套完整方案，评委先选择其中一个方案（落选方案直接PASS），再对选中方案的每个决策点逐项打分（0-100分制），加权平均计算总分\n- **决策点拆分评审**：复杂/多维度方案，组织者将方案拆解为若干「决策点」，每个决策点对应一个具体评审维度。评委对每个决策点逐项打分（0-100分制）。整体方案得分为各决策点评分的加权平均值，由组织者自动计算，不再由评委单独打分\n\n---\n\n### 第 1 轮：独立评估 (Blind Evaluation)\n\n**动作说明**：组织者接收到决策确认后，将准备流程相关内容分发至各选定模型的独立子会话中。每位委员在**完全无法感知他方意见**的情况下进行背对背评审。\n\n**评审维度**：方案逻辑合理性、完整性、潜在风险分析、实施资源消耗预估、可行性评审、优化建议，并按照100分值进行评分。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：组织者汇总各委员的初评报告及投票结果及打分结果，以及分歧意见或争议点。\n\n---\n\n### 第 2 轮：分歧讨论 (Dispute Discussion)\n\n**动作说明**：组织者将第1轮中标记为「未通过」的决策点整理为「未通过决策点清单」，仅将这些未通过项下发给所有委员，要求评委**只针对这些未通过决策点**进行讨论并重新评分。已通过的决策点不再讨论。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：汇总本轮重新评分结果，若所有决策点全员通过则直接输出最终报告；若仍有未通过项则进入下一轮。\n\n---\n\n### 最后一轮：最终辩论 (Final Duel)\n\n**特殊说明**：无论配置多少轮，最终辩论始终是**最后一轮**。\n- 3轮配置：第3轮为最终辩论\n- 5轮配置：第5轮为最终辩论\n- N轮配置：第N轮为最终辩论\n\n**动作说明**：经过分歧讨论后，仍有决策点未通过时启用。组织者将仍未通过的决策点再次下发给所有委员，要求评委进行最后一轮讨论和重新评分。\n\n**评审时长**：不超过3分钟/轮。若3分钟仍未回收所有子会话，组织者将结束流程并返回结果。\n\n**核心产出**：若所有决策点全员通过，输出最终报告；若仍有未通过项，将未通过项标注状态后输出最终报告。\n\n---\n\n### 轮次扩展说明\n\n| 扩展类型 | 说明 |\n|:---|:---|\n| 增加轮次（4轮及以上） | 倒数第2轮参照「分歧讨论」；最后一轮参照「最终辩论」；中间轮次循环分歧讨论 |\n| 减少轮次（2轮） | 第1轮参照\"独立评估\"，第2轮直接输出最终报告 |\n\n---\n\n## 最终输出：6段式共识报告\n\n决策完成后，输出标准化报告，包含：\n\n1. **投票记录** — 量化评分矩阵（含决策点级别评分）\n2. **汇总说明** — 各委员核心论点（注明大模型）\n3. **执行方案** — 结论版 + 说明版\n4. **决策点通过清单** — 各决策点通过状态\n5. **结论摘要** — 一句话裁定 + 置信度\n6. **风险提示** — 主要风险与缓解建议\n\n**决策点通过状态标注**：\n\n| 状态 | 含义 | 标注 |\n|:---|:---|:---|\n| 全员通过 | 所有评委在该决策点评分均≥阈值 | ✅ 绿色 — 通过 |\n| 待决策 | 推荐顺序一致但有评委评分低于阈值 | 🟡 黄色 — 待用户确认 |\n| 有分歧 | 推荐顺序不一致 | 🔴 红色 — 有分歧，附原因 |\n\n---\n*版本: V1.6.7*\n*开发者: Zeekr0808*\n*邮箱: Zeekr0808@outlook.com*\n\nFile v1.7.1:_meta.json\n\n{\n  \"ownerId\": \"kn7e3gt0d3qkc3nmkhqzcx15ks85fz9s\",\n  \"slug\": \"multi-model-consensus\",\n  \"version\": \"1.7.1\",\n  \"publishedAt\": 1778816649578\n}\n\nFile v1.7.1:references/OUTPUT_TEMPLATE.md\n\n# 多模型决策委员会 — 输出模板\n\n## 6段式共识报告模板\n\n---\n\n# 多模型决策共识报告\n\n**决策主题**：（用户输入的决策任务简述）\n**生成时间**：（自动生成时间戳）\n**评委配置**：（模型组合，如 Kimi K2.6 + GPT-5.4-mini + GLM-5.1）\n**决策轮数**：（实际执行轮数）\n**通过阈值**：（如 ≥90%）\n\n---\n\n## 一、投票记录\n\n| 评委 | 方案A评分 | 方案B评分 | 推荐顺序 | 核心论点 |\n|:---|:---:|:---:|:---:|:---|\n| Kimi K2.6（实名） | X/100 | X/100 | A > B | （简述） |\n| GPT-5.4-mini（实名） | X/100 | X/100 | B > A | （简述） |\n| GLM-5.1（实名） | X/100 | X/100 | A > B | （简述） |\n| **加权汇总** | **X/100** | **X/100** | **A优先** | 置信度：XX% |\n\n---\n\n## 二、汇总说明\n\n### Kimi K2.6（实名） — 核心论点\n\n（详细说明该评委的分析逻辑、评分依据、主要担忧）\n\n### GPT-5.4-mini（实名） — 核心论点\n\n（详细说明该评委的分析逻辑、评分依据、主要担忧）\n\n### GLM-5.1（实名） — 核心论点\n\n（详细说明该评委的分析逻辑、评分依据、主要担忧）\n\n---\n\n## 三、执行方案\n\n### 结论版（一句话）\n\n> 建议采用方案A，主要理由为：（最核心的1-2个理由）\n\n### 详细版\n\n**适用场景**：（方案A的推荐条件）\n**预期收益**：（量化收益或质化收益）\n**必要前提**：（必须满足的前置条件）\n**执行步骤**：\n1. （第一步）\n2. （第二步）\n3. （第三步）\n\n---\n\n## 四、未决清单\n\n| 事项 | 负责方 | 完成时间 | 优先级 |\n|:---|:---|:---:|:---:|\n| （待确认事项1） | 负责人 | YYYY-MM-DD | 高 |\n| （待确认事项2） | 负责人 | YYYY-MM-DD | 中 |\n\n---\n\n## 五、结论摘要\n\n> **一句话裁定**：建议（不）采用方案A\n> **置信度**：XX%（基于X轮评审，X/X评委推荐）\n> **主要风险**：XX风险需重点关注\n\n---\n\n## 六、风险提示\n\n| 风险类型 | 风险描述 | 概率 | 影响 | 缓解措施 |\n|:---|:---|:---:|:---:|:---|\n| 技术风险 | （描述） | 中 | 高 | （措施） |\n| 进度风险 | （描述） | 低 | 中 | （措施） |\n| 资源风险 | （描述） | 中 | 中 | （措施） |\n\n---\n\n## 附录：决策过程回顾\n\n- **Round 1 完成时间**：（时间戳）\n- **Round 2 完成时间**：（时间戳）\n- **倒数第2轮完成时间**：（时间戳）\n- **最后一轮（最终辩论）完成时间**：（时间戳，若有执行）\n- **未通过决策点**：（第1轮结束后仍有哪些决策点未达全员通过）\n- **第2轮分歧讨论结果**：（第2轮讨论后是否所有决策点全员通过）\n- **异常记录**：（如有评委超时或失败，在此注明）\n\n---\n\n## 第1轮独立评审 Prompt 模板\n\n```\n【角色】你是多模型决策委员会的评委委员，你的身份是：{标准大模型名称}（实名委员）。\n【任务】独立评审以下方案，不与任何其他评委交流。\n       【硬性约束】评委严禁spawn子Agent或调用任何外部工具，必须基于自身知识独立完成评审。\n【评审方案】\n（用户提交的方案内容）\n\n【评审维度】（6维度，每维度100分制）\n1. 逻辑完整性：（方案推理是否严密）— X/100\n2. 风险控制：（潜在风险识别与应对）— X/100\n3. 实施资源：（成本、人力、时间估算）— X/100\n4. 可行性：（技术与业务可行性）— X/100\n5. 长期价值：（战略意义与可持续性）— X/100\n6. 用户接受度：（用户体验与采用意愿）— X/100\n【硬约束】你只能输出一套评分。\n         禁止在同一输出中出现两套或以上相互独立的评分体系。\n         若涉及多个视角，请在「主要担忧」中简要注明，不要拆分为多个独立评分。\n\n【输出格式】\n决策点评分：[每个决策点0-100分，列出所有决策点]\n推荐顺序：[方案A优先 / 方案B优先 / 持平]\n核心论点：[200字以内的主要理由]\n主要担忧：[100字以内的主要风险点]\n```\n\n---\n\n## 第2轮分歧讨论 Prompt 模板\n\n```\n【角色】你是多模型决策委员会的评委委员（实名：{标准大模型名称}）。\n【背景】第1轮独立评审已完成，以下是各评委对各决策点的评分结果（实名）：\n（汇总第1轮决策点评分矩阵，注明各评委真实模型名称）\n\n【未通过决策点清单】\n（列出第1轮中未达到全员通过的决策点，即：任一评委评分<阈值的决策点）\n\n【任务】请阅读其他评委的观点后，**仅针对以上未通过决策点**进行二次表态和重新评分。\n- 若同意某评委的观点，请说明理由\n- 若坚持自己的观点，请提供更强论据\n- 若改变立场，明确说明原因\n\n【输出格式】\n对分歧点1的态度：[坚持/改变/部分同意] — 理由（50字以内）\n对分歧点2的态度：[坚持/改变/部分同意] — 理由（50字以内）\n决策点重新评分：[每个未通过决策点重新打分，0-100分]\n最终推荐顺序：[是否与第1轮一致]\n```\n\n---\n\n## 最后一轮最终辩论 Prompt 模板\n\n> **特殊说明**：无论配置多少轮，最终辩论始终是最后一轮。\n\n```\n【角色】你是多模型决策委员会的评委委员（实名：{标准大模型名称}）。\n【背景】经过第2轮分歧讨论后，仍存在以下未通过的决策点（实名评委）：\n（列出仍未解决的分歧点，标注各评委真实身份）\n\n【任务】针对以上分歧，将其简化为二元对立选项（是/否 或 方案A/方案B），给出你最终的逻辑立场。\n\n【通过条件】所有评委在所有决策点上的评分均≥阈值 = 通过；若有决策点未通过，标注其状态（✅/🟡/🔴）。\n\n【输出格式】\n对分歧点1的最终立场：[选项A / 选项B] — 理由（50字以内）\n对分歧点2的最终立场：[选项A / 选项B] — 理由（50字以内）\n最终推荐顺序：[方案A优先 / 方案B优先 / 持平]\n（达到推荐顺序一致时即视为通过）\n```\n\nFile v1.7.1:references/SCHEMA.md\n\n# 多模型决策委员会 — 状态文件 Schema（已归档）\n\n> ⚠️ **V1.6.0 说明**：本文档描述的状态文件 Schema 已归档。V1.6.0 起不再使用持久化状态文件，改用「轻量级状态跟踪」（会话上下文变量）。本文件保留供历史参考。\n\n## 状态文件字段\n\n| 字段 | 类型 | 必填 | 说明 |\n|:---|:---|:---:|:---|\n| `state` | string | ✅ | 当前状态：IDLE / ROUND_0 / ROUND_1 / ROUND_2 / ROUND_3 / COMPLETE |\n| `config` | object | ✅ | 评委配置 |\n| `config.models` | array | ✅ | 评委模型列表，如 [\"模型A\", \"模型B\", \"模型C\"] |\n| `config.rounds` | integer | ✅ | 决策轮数，2-6，默认3 |\n| `config.threshold` | number | ✅ | 通过阈值，0-1，如 0.9 |\n| `round0` | object | — | 准备阶段汇总内容 |\n| `round1` | object | — | 第1轮结果（含 judges[].used_subagent） |\n| `round2` | object | — | 第2轮结果 |\n| `round3` | object | — | 第3轮结果（仅在有分歧时） |\n| `report` | object | — | 最终6段式共识报告 |\n\n---\n\n## 版本演进规则\n\n1. **不得单方面升版**：任何 Schema 变更必须经过委员会评审通过\n2. **向后兼容**：新版本应尽可能兼容旧版本数据格式\n3. **升版记录**：每次升版须在 CHANGELOG.md 中详细记录变更内容\n\nFile v1.7.1:references/STATE_MACHINE.md\n\n# 多模型决策委员会 — 状态机规范（概念参考）\n\n> ⚠️ **V1.6.0 说明**：本文档描述的状态机逻辑仍为概念规范，但实际执行时已**不再使用持久化状态文件**。组织者改用「轻量级状态跟踪」（会话上下文变量）来跟踪评审进度。详见 SKILL.md 「轻量级状态跟踪」章节。\n> \n> 本文件保留作为状态流转的概念参考和异常处理指南。\n\n## 状态定义（6状态）\n\n| 状态 | 说明 | 可执行动作 |\n|:---|:---|:---|\n| `IDLE` | 初始状态，无进行中决策 | 接受任务 → ROUND_0 |\n| `ROUND_0` | 准备阶段，组织者汇总决策申请内容 | 分发方案 → ROUND_1 |\n| `ROUND_1` | 第1轮独立评审中 | 汇总报告 → ROUND_2 |\n| `ROUND_2` | 第2轮收敛讨论中 | 判断收敛性 → ROUND_3 或 COMPLETE |\n| `ROUND_3` | 第3轮最终辩论中（仅在ROUND_2有分歧时触发） | 汇总矩阵 → COMPLETE |\n| `COMPLETE` | 决策完成，报告已生成 | 输出报告 → IDLE |\n\n---\n\n## 状态流转图\n\n```\n[IDLE] → 用户触发 → [ROUND_0]\n                         ↓\n                  汇总决策申请内容\n                         ↓\n[ROUND_1] ← 分发方案 → [ROUND_0]\n     ↓\n  有分歧          无分歧\n     ↓              ↓\n[ROUND_2]      [COMPLETE]\n     ↓\n  有分歧          无分歧\n     ↓              ↓\n[ROUND_3]      [COMPLETE]\n     ↓              ↓\n  汇总矩阵    输出6段式报告\n     ↓              ↓\n[COMPLETE] ───────→ [IDLE]\n```\n\n---\n\n## 收敛判定规则\n\n### 收敛条件\n\n满足以下**全部条件**时，判定为收敛，跳过下一轮：\n\n1. 所有评委的**推荐顺序完全一致**\n2. 无新产生的争议点\n3. 所有评委的综合评分差距在阈值范围内\n\n### 分歧条件\n\n满足以下**任一条件**时，判定为有分歧，触发下一轮：\n\n1. 任意两位评委的**推荐顺序冲突**\n2. 存在**未解决的反对意见**\n3. 评委评分差距超过阈值范围（>20%）\n\n---\n\n## 异常处理\n\n| 异常情况 | 处理方式 |\n|:---|:---|\n| 某评委子会话超时（3分钟） | 报告中标注\"评委X超时\"，不影响其他评委结果 |\n| 某评委返回格式错误 | 使用前一轮结果替代，报告中注明 |\n| 用户中断决策 | 保留当前状态文件，下次可恢复执行 |\n\n---\n\n## 状态文件结构\n\n```json\n{\n  \"state\": \"ROUND_1\",\n  \"config\": {\n    \"models\": [\"模型A\", \"模型B\", \"模型C\"],\n    \"rounds\": 3,\n    \"threshold\": 0.9\n  },\n  \"round0\": {\n    \"completed_at\": \"2026-04-25T00:00:00Z\",\n    \"preparation_content\": {\n      \"项目名称\": \"...\",\n      \"项目背景\": \"...\",\n      \"项目需求\": \"...\",\n      \"待审方案\": \"...\"\n    }\n  },\n  \"round1\": {\n    \"completed_at\": \"2026-04-25T00:01:00Z\",\n    \"judges\": [\n      {\n        \"model\": \"模型A\",\n        \"used_subagent\": false,\n        \"result\": { \"score\": 85, \"recommendation\": \"A > B\", \"arguments\": \"...\" }\n      }\n    ]\n  },\n  \"round2\": {\n    \"completed_at\": \"2026-04-25T00:02:00Z\",\n    \"disputes\": []\n  }\n}\n```\n\nFile v1.7.1:references/TROUBLESHOOTING.md\n\n# 多模型决策委员会 — 常见失败模式与排查指南\n\n> 版本：V1.6.0  \n> 本文档记录自动化执行中的常见失败模式及排查方法。\n\n---\n\n## 失败模式 1：Thread binding invalid\n\n**症状**：\n```\nerrorCode: \"thread_binding_invalid\"\nerror: \"Thread bindings are unavailable f\n\nArchive v1.7.0: 10 files, 54149 bytes\n\nFiles: _meta.json (140b), docs/README.md (24492b), docs/USER_GUIDE.md (39514b), README.md (24445b), references/OUTPUT_TEMPLATE.md (6104b), references/SCHEMA.md (1314b), references/STATE_MACHINE.md (3054b), references/TROUBLESHOOTING.md (5964b), references/VERIFICATION_CASE.md (4342b), SKILL.md (19246b)\n\nArchive v1.9.1: 11 files, 55066 bytes\n\nFiles: _meta.json (140b), docs/README.md (24492b), docs/USER_GUIDE.md (39514b), README.md (24445b), references/OUTPUT_TEMPLATE.md (6104b), references/SCHEMA.md (1314b), references/STATE_MACHINE.md (3054b), references/TROUBLESHOOTING.md (5964b), references/VERIFICATION_CASE.md (4342b), skill-card.md (2398b), SKILL.md (18005b)\n\nArchive v1.6.9: 10 files, 53703 bytes\n\nFiles: _meta.json (140b), docs/README.md (24492b), docs/USER_GUIDE.md (39514b), README.md (24445b), references/OUTPUT_TEMPLATE.md (6104b), references/SCHEMA.md (1314b), references/STATE_MACHINE.md (3054b), references/TROUBLESHOOTING.md (5964b), references/VERIFICATION_CASE.md (4342b), SKILL.md (18005b)\n\nArchive v1.6.8: 10 files, 53704 bytes\n\nFiles: _meta.json (140b), docs/README.md (24492b), docs/USER_GUIDE.md (39514b), README.md (24445b), references/OUTPUT_TEMPLATE.md (6104b), references/SCHEMA.md (1314b), references/STATE_MACHINE.md (3054b), references/TROUBLESHOOTING.md (5964b), references/VERIFICATION_CASE.md (4342b), SKILL.md (18005b)\n\nArchive v1.9.0: 10 files, 53702 bytes\n\nFiles: _meta.json (140b), docs/README.md (24492b), docs/USER_GUIDE.md (39514b), README.md (24445b), references/OUTPUT_TEMPLATE.md (6104b), references/SCHEMA.md (1314b), references/STATE_MACHINE.md (3054b), references/TROUBLESHOOTING.md (5964b), references/VERIFICATION_CASE.md (4342b), SKILL.md (18005b)\n\nArchive v1.6.7: 10 files, 53686 bytes\n\nFiles: _meta.json (140b), docs/README.md (24492b), docs/USER_GUIDE.md (39514b), README.md (24445b), references/OUTPUT_TEMPLATE.md (6104b), references/SCHEMA.md (1314b), references/STATE_MACHINE.md (3054b), references/TROUBLESHOOTING.md (5964b), references/VERIFICATION_CASE.md (4342b), SKILL.md (17975b)\n\nArchive v1.6.6: 10 files, 53140 bytes\n\nFiles: _meta.json (140b), docs/README.md (24221b), docs/USER_GUIDE.md (38923b), README.md (23875b), references/OUTPUT_TEMPLATE.md (6152b), references/SCHEMA.md (1314b), references/STATE_MACHINE.md (3054b), references/TROUBLESHOOTING.md (5964b), references/VERIFICATION_CASE.md (4342b), SKILL.md (17954b)\n\nArchive v1.6.5: 10 files, 52491 bytes\n\nFiles: _meta.json (140b), docs/README.md (24008b), docs/USER_GUIDE.md (38410b), README.md (23393b), references/OUTPUT_TEMPLATE.md (6152b), references/SCHEMA.md (1314b), references/STATE_MACHINE.md (3054b), references/TROUBLESHOOTING.md (5964b), references/VERIFICATION_CASE.md (4363b), SKILL.md (17495b)","readmeExcerpt":"Skill: Multi Model Consensus Owner: zeekr0808-hue Summary: 多模型决策委员会 — 消除单模型偏见，通过多轮分歧讨论产出客观决策参考。支持3-6个模型同时评审，提供量化投票矩阵和6段式共识报告。触发条件：包含「多模型决策」或「多模型委员会」时自动激活。 Tags: latest:1.9.1 Version history: v1.8.0 | 2026-05-16T14:47:53.996Z | auto **Summary:** This release enforces stricter process control and standardization for the multi-model decision committee, focusing on compliance, template usage, and evaluation transparency.","codeSnippets":[],"executableExamples":[{"language":"text","snippet":"Spawn 子Agent → sessions_yield 挂起 → 收到 subagent_announce 事件 → 提取结果 → 更新跟踪"},{"language":"yaml","snippet":"SessionsSpawn:\n  runtime: \"subagent\"\n  mode: \"run\"\n  model: \"n1n/gpt-5.4\"\n  task: \"评审任务...\"\n  timeoutSeconds: 180"},{"language":"yaml","snippet":"SessionsSpawn:\n  runtime: \"acp\"\n  mode: \"session\"\n  thread: true\n  streamTo: \"parent\"\n  model: \"n1n/gpt-5.4\"\n  task: \"评审任务...\"\n  timeoutSeconds: 180"},{"language":"text","snippet":"[本轮状态跟踪]\nRound: {1/2/3}\nSpawned: {N}\nReceived: {M} ({评委A} ✅, {评委B} ✅)\nPending: {N-M} ({评委C} ⏳)"},{"language":"text","snippet":"═══════════════════════════════════════════════\n🏛️ 多模型决策委员会 — 决策准备确认\n═══════════════════════════════════════════════\n\n【项目名称】\n（从用户输入提炼，不超过20字）\n\n【项目背景】\n（描述当前情境、痛点或决策缘由，50字以内）\n\n【项目需求】\n（用户期望解决的核心问题，50字以内）\n\n【待审方案】\n（完整呈现待评审方案内容）\n\n───────────────────────────────────────────────\n【评审规则】：各位评委背对背独立评审，由组织者汇总。若所有评委评审项均达到通过阈值，该项视为**通过**；若所有评委评审项均未达到否定阈值，该项视为**不通过**；其余情况视为**待定**。\n\n【决策点拆分及权重】\n  决策点1（{维度名称}）：{权重}%\n  决策点2（{维度名称}）：{权重}%\n  ……\n  （权重总和须等于100%）\n\n───────────────────────────────────────────────\n【评委配置】\n| 序号 | 委员 | 模型 | 备注 |\n|:---:|:---|:---|:---|\n| 1 | 评委1 | {模型名}（{Provider}） | |\n| 2 | 评委2 | {模型名}（{Provider}） | |\n| 3 | 评委3 | {模型名}（{Provider}） | |\n\n【决策轮次】：{3} 轮\n【阈值设置】：通过≥90% / 待定[60%,90%) / 否定<60%\n【判定方式】：全票通过\n\n是否需要修改「评委」、「决策轮次」、「通过阈值」、「判定方式」等配置？若有需要，请告诉我。\n\n═══════════════════════════════════════════════\n请确认以上框架后，回复「确认」或「开始决策」。\n═══════════════════════════════════════════════"},{"language":"text","snippet":"═══════════════════════════════════════════════\n📊 第 1 轮独立评估 — 评分汇总\n═══════════════════════════════════════════════\n\n【评审类型】：决策点拆分评审\n【评委到齐】：{X}/{X}（{缺席评委名} 超时未提交）\n\n───────────────────────────────────────────────\n【本轮判定结果】：{通过 / 待定 / 不通过}\n阈值设置：通过≥90% / 待定[60%,90%) / 否定<60%\n═══════════════════════════════════════════════"}],"parameters":null,"dependencies":[],"permissions":[],"extractedFiles":[{"path":"SKILL.md","content":"---\nname: multi-model-consensus\ndescription: 多模型决策委员会 — 消除单模型偏见，通过多轮分歧讨论产出客观决策参考。支持3-6个模型同时评审，提供量化投票矩阵和5段式共识报告。触发条件：包含「多模型决策」或「多模型委员会」时自动激活。\nallowed-tools: SessionsSpawn,SessionsSend,SessionsHistory,Read,Write\nmetadata:\n  openclaw:\n    emoji: \"🏛️\"\n    requires:\n      capability: sessions_spawn\n    # 路径约束：见正文\"技术约束\"章节\n---\n\n# 🏛️ 多模型决策委员会\n\n专为 OpenClaw 设计的\"数字智库\"，通过多模型独立评审与共识合成，消除单模型偏见，为复杂决策提供客观参考。\n\n---\n\n\n## 🎯 插件定位\n\n**多模型决策委员会** 是专为 OpenClaw 设计的\"数字智库\"。它支持调动多个不同架构的大模型（如 GPT、Gemini、豆包等）同时对同一任务进行协同思考、对抗辩论与共识合成，有效消除单模型偏见，综合性给出最优建议。\n\n---\n\n## 核心优势\n\n- **去中心化决策**：交叉验证不同模型逻辑，确保方案的严谨性。\n- **对抗性评审**：通过模型间的互评，快速锁定潜在的逻辑漏洞。\n- **弹性扩容**：根据任务难度，随时增加或减少\"决策委员\"的数量。\n- **消除偏见**：多模型独立评审，避免单一AI的认知盲区。\n- **量化决策**：6维度评分 + 加权矩阵，结论有据可查。\n- **透明可信**：实名委员制、模型身份透明。\n- **灵活配置**：参数可调，如决策成员人数、决策轮次、通过阈值等，均可可自定义。\n\n---\n\n## 核心治理原则：身份纯净\n\n为了确保结果的绝对客观，本插件强制执行 **\"身份纯净原则\"**：\n- **角色透明化**：决策委员会成员均会注明其身份（大模型名称），保持可视化透明。\n- **严禁设定角色**：禁止给委员设定诸如\"架构师\"、\"审计员\"等身份标签。\n- **严禁引导提示**：任务下发时不得针对评审委员设置诱导性提示词或诱导性身份角色，防止产生的偏见。\n- **评委工具隔离**：评委子Agent仅通过 `task` 参数接收评审内容，独立输出结论；不得调用任何工具（包括 spawn、Read、Write、Exec、Search 等），所有外部操作由组织者统一管理\n- **评审不重复原则**：若组织者（即当前使用模型）是评审委员会成员，则组织者提交内容即视同为其在本轮的评审结果，无需重复自评。若组织者不是评审委员会成员，则仅负责组织实施和汇总评委意见，不参与评审。\n\n---\n\n## 使用方法\n\n- **启动**：在OpenClaw中，当用户输入的内容包含「多模型决策」或「多模型委员会」时，自动激活多模型决策委员会。首次使用时会提醒用户进入配置模式，并选择模型组合。\n- **配置**：根据需要，可配置参数，如：轮数、阈值、委员数量、委员模型等。\n\n### 参数配置\n\n配置参数可采用指令方式，如：\"修改配置\"、\"调整参数\"、\"增加委员\"、\"确认配置\"。\n\n| 指令关键词 | 可调参数 | 说明 |\n|:---|:---|:---|\n| \"换模型\" | 委员模型组合 | 从用户本地接入的模型列表中选择模型组合加入。用户要求调整评委时，组织者**必须**读取 `openclaw.json` 列出当前环境下所有可接入的大模型名单（含Provider、上下文、最大输出等关键信息），供用户选择。 |\n| \"改轮数\" | 决策轮数 | 默认3轮，日常事务可设为2轮，重大决策可设3-6轮 |\n| ~~\"收敛分差阈值\"~~ | ~~收敛判定分差~~ | ~~已废弃~~ |\n| \"改阈值\" | 判定阈值 | **通过阈值**：默认≥90%；**待定阈值**：默认[60%, 90%)；**否定阈值**：默认<60% |\n| \"改判定方式\" | 通过判定方式 | **全票通过** |\n| \"增加委员\" / \"减少委员\" | 委员数量 | 支持2-6个模型同时评审。用户要求增减委员时，组织者**必须**先列出当前环境下所有可接入的大模型名单，供用户指定。 |\n| \"确认配置\" | 确认当前配置 | 确认当前参数后开始决策 |\n\n### 评委调整规则（铁律）\n\n当用户表达\"换评委\"、\"换模型\"、\"调整评委\"、\"增减委员\"等意图时，组织者**必须**执行以下步骤：\n\n1. **列出全部可用模型**：读取 `openclaw.json` 中所有 provider 的 models 列表，完整展示：\n   - 模型ID（如 `gpt-5.5`、`gemini-3.1-pro-preview`）\n   - 模型名称（如 GPT-5.5、Gemini 3.1 Pro）\n   - Provider来源（如 n1n、n1n-gemini、ark-cn）\n   - 上下文窗口（contextWindow）\n   - 最大输出（maxTokens）\n2. **推荐组合**：基于任务特点，给出1-3组推荐搭配（如\"逻辑+中文\"、\"深度+性价比\"）。\n3. **等待用户指定**：禁止直接替用户选择，必须等待用户明确指定模型组合后方可继续。\n\n### 参数建议\n\n- **模型选择**：建议至少包含 1 个具备强逻辑推理能力的模型和 1 个具备强中文语境理解能力的模型。\n- **轮次建议**：日常事务建议 2 轮；涉及架构，资金、核心规则的决策，建议 3-5 轮。\n- **阈值设定**：\n  - **严谨型**：≥95%（全票通过）\n  - **效率型**：≥75%（多数通过）\n\n---\n\n## 决策流程（3轮决策机制）\n\n多模型决策委员会采用 3 轮决策机制，基于决策点级别评审，每轮进行100分制评分，最终给出决策结果。\n\n---\n\n### 第 0 轮：准备与框架确认 (Preparation & Framework Confirmation)\n\n**决策准备**：组织者（当前使用模型）接收到决策任务后，将用户背景、需求、待审方案汇总，拆解为决策点及权重，发起决策申请。\n\n> 📌 **标准模板**：组织者须完整套用 `references/OUTPUT_TEMPLATE.md` 中的 **T1 决策准备确认模板**，不得自行增删字段或自由发挥。\n\n---\n\n### 运行时自检（第 0 轮环境兼容性检查）\n\n组织者在启动评委前，执行环境自检以确保子Agent可正常 spawn 和返回结果。此步骤为系统自动执行，确保运行环境兼容性。\n\n**自检方法**：通过 `session_status` 或检查当前会话元数据，确认通道类型。\n\n**环境适配矩阵**：\n\n| 环境 | Runtime | Mode | 额外参数 | 结果回收方式 |\n|:---|:---|:---|:---|:---|\n| **Webchat** | `subagent` | `\"run\"` | 无 | `subag"},{"path":"docs/README.md","content":"# Multi-Model Consensus Council — Operation & User Guide\n\n## 📋 Version History\n\n| Version | Date | Changes |\n|:---|:---|:---|\n| V1.0.0 | 2026-04-19 | Initial release: core framework, 3-round convergence, 6-section report |\n| V1.1.0 | 2026-04-24 | Added English operation guide (merged with Chinese, English first); translation reviewed and approved by multi-model committee (Gemini/Doubao/GLM, avg 81/100); wording optimizations |\n| V1.1.3 | 2026-04-24 | SKILL.md and references/ converted to pure Chinese for AI readability; docs/ retains full bilingual documentation for global community |\n| V1.2.0 | 2026-04-25 | Restored public ClawHub version (was local customized); sync all files |\n| V1.2.1 | 2026-04-25 | **CRITICAL FIX**: Removed hardcoded model list (A1/A2/A4...) from SKILL.md; replaced with dynamic scan via `openclaw models list`; examples updated to use generic `[Model A/B/C]` placeholders |\n| V1.2.2 | 2026-04-25 | SYNC: Full file sync to GitHub and ClawHub after v1.2.1 hotfix |\n| V1.2.3 | 2026-04-25 | **UX FIX**: Clarify first-time user flow: auto-scan models → default to first 3 → prompt user to confirm/modify before starting decision |\n| V1.2.4 | 2026-04-25 | **CRITICAL FIX**: OUTPUT_TEMPLATE.md - remove \"匿名委员\" from all 3 round prompt templates; replaced with real-name format `{模型名称}（实名委员）` |\n| V1.2.5 | 2026-04-25 | **ENFORCEMENT**: SKILL.md - add mandatory threshold check rules; add forbidden items for skipping threshold/convergence checks; state machine now has explicit checkpoint enforcement |\n| V1.2.6 | 2026-04-25 | **Real-name Transparency**: Judges use standard model names; 3-layer architecture (organizer/judge/sub-agent); 100-point scoring; Round 0 preparation phase added |\n| V1.5.0 | 2026-04-25 | Sync all document versions to V1.5.0; renamed SCHEMA.md; full English content aligned with Chinese |\n| V1.5.1 | 2026-04-26 | Added round-start user notification rule; added convergence score threshold parameter; added exception handling rules section |\n| V1.5.2 | 2026-04-26 | Added judgment method configurable parameter; clarified unanimous-pass rule as default |\n\n---\n\n## 🎯 Plugin Overview\n\n**Multi-Model Consensus Council** is a \"Digital Think Tank\" designed for OpenClaw. It coordinates multiple large language models (e.g., GPT, Gemini, Doubao) to collaboratively analyze, debate, and synthesize decisions, effectively eliminating single-model bias.\n\n### Core Advantages\n\n- **Decentralized Decision-Making**: Cross-validates reasoning across models, ensuring rigor.\n- **Adversarial Review**: Rapidly identifies logical flaws through peer evaluation.\n- **Elastic Scaling**: Adjust the number of \"decision committee members\" anytime based on task complexity.\n- **Bias Elimination**: Multi-model independent review avoids single-AI cognitive blind spots.\n- **Quantitative Decision-Making**: 6-dimension scoring + weighted matrix, conclusions are evidence-based.\n- **Transparency & Trust**: Real-name committee system, model identity visible.\n- **Flexible C"},{"path":"README.md","content":"# Multi-Model Consensus Council — Operation & User Guide\n\n## 📋 Version History\n\n| Version | Date | Changes |\n|:---|:---|:---|\n| V1.0.0 | 2026-04-19 | Initial release: core framework, 3-round convergence, 6-section report |\n| V1.1.0 | 2026-04-24 | Added English operation guide (merged with Chinese, English first); translation reviewed and approved by multi-model committee (Gemini/Doubao/GLM, avg 81/100); wording optimizations |\n| V1.1.3 | 2026-04-24 | SKILL.md and references/ converted to pure Chinese for AI readability; docs/ retains full bilingual documentation for global community |\n| V1.2.0 | 2026-04-25 | Restored public ClawHub version (was local customized); sync all files |\n| V1.2.1 | 2026-04-25 | **CRITICAL FIX**: Removed hardcoded model list (A1/A2/A4...) from SKILL.md; replaced with dynamic scan via `openclaw models list`; examples updated to use generic `[Model A/B/C]` placeholders |\n| V1.2.2 | 2026-04-25 | SYNC: Full file sync to GitHub and ClawHub after v1.2.1 hotfix |\n| V1.2.3 | 2026-04-25 | **UX FIX**: Clarify first-time user flow: auto-scan models → default to first 3 → prompt user to confirm/modify before starting decision |\n| V1.2.4 | 2026-04-25 | **CRITICAL FIX**: OUTPUT_TEMPLATE.md - remove \"匿名委员\" from all 3 round prompt templates; replaced with real-name format `{模型名称}（实名委员）` |\n| V1.2.5 | 2026-04-25 | **ENFORCEMENT**: SKILL.md - add mandatory threshold check rules; add forbidden items for skipping threshold/convergence checks; state machine now has explicit checkpoint enforcement |\n| V1.2.6 | 2026-04-25 | **Real-name Transparency**: Judges use standard model names; 3-layer architecture (organizer/judge/sub-agent); 100-point scoring; Round 0 preparation phase added |\n| V1.5.0 | 2026-04-25 | Sync all document versions to V1.5.0; renamed SCHEMA.md; full English content aligned with Chinese |\n| V1.5.1 | 2026-04-26 | Added round-start user notification rule; added convergence score threshold parameter; added exception handling rules section |\n| V1.5.2 | 2026-04-26 | Added judgment method configurable parameter; clarified unanimous-pass rule as default |\n| **V1.6.0** | **2026-04-27** | **MAJOR FIX**: Rewrote sub-agent result recovery; added runtime environment adaptation matrix; Push-based waiting flow; runtime self-check; simplified state management; added TROUBLESHOOTING.md |\n| **V1.6.1** | **2026-04-27** | **Clarified judgment rules**: Separated \"Pass\" vs \"Convergence\" vs \"Consensus Selection\"; clarified convergence discussion scope; renamed \"Round 3\" to \"Final Round\"; added Consensus Selection for final round |\n| **V1.6.2** | **2026-04-27** | **Core logic refactored**: Removed \"Convergence\"; introduced \"Decision Point\" level review; \"Convergence Discussion\" renamed to \"Dispute Discussion\"; final report includes decision point status (✅/🟡/🔴) |\n| **V1.6.5** | **2026-04-30** | **Security Hardening**: Removed unused `Exec` from allowed-tools; reduced max concurrent sub-agents from 13 to 6; added `Write` tool path restriction (mem"},{"path":"_meta.json","content":"{\n  \"ownerId\": \"kn7e3gt0d3qkc3nmkhqzcx15ks85fz9s\",\n  \"slug\": \"multi-model-consensus\",\n  \"version\": \"1.8.0\",\n  \"publishedAt\": 1778942873996\n}"},{"path":"references/OUTPUT_TEMPLATE.md","content":"# 多模型决策委员会 — 输出模板\n\n> 本文件定义各阶段标准模板。T2/T4/T6 为评委分发 Prompt，已在末尾附录；T1/T3/T5/T7 为本文件正文模板。\n\n---\n\n## T1 — 决策准备确认模板（第0轮）\n\n组织者向用户确认决策框架，**须完整填写以下所有字段**：\n\n```\n═══════════════════════════════════════════════\n🏛️ 多模型决策委员会 — 决策准备确认\n═══════════════════════════════════════════════\n\n【项目名称】\n（从用户输入提炼，不超过20字）\n\n【项目背景】\n（描述当前情境、痛点或决策缘由，50字以内）\n\n【项目需求】\n（用户期望解决的核心问题，50字以内）\n\n【待审方案】\n（完整呈现待评审方案内容）\n\n───────────────────────────────────────────────\n【评审规则】：各位评委背对背独立评审，由组织者汇总。若所有评委评审项均达到通过阈值，该项视为**通过**；若所有评委评审项均未达到否定阈值，该项视为**不通过**；其余情况视为**待定**。\n\n【决策点拆分及权重】\n  决策点1（{维度名称}）：{权重}%\n  决策点2（{维度名称}）：{权重}%\n  ……\n  （权重总和须等于100%）\n\n───────────────────────────────────────────────\n【评委配置】\n| 序号 | 委员 | 模型 | 备注 |\n|:---:|:---|:---|:---|\n| 1 | 评委1 | {模型名}（{Provider}） | |\n| 2 | 评委2 | {模型名}（{Provider}） | |\n| 3 | 评委3 | {模型名}（{Provider}） | |\n\n【决策轮次】：{3} 轮\n【阈值设置】：通过≥90% / 待定[60%,90%) / 否定<60%\n【判定方式】：全票通过\n\n是否需要修改「评委」、「决策轮次」、「通过阈值」、「判定方式」等配置？若有需要，请告诉我。\n\n═══════════════════════════════════════════════\n请确认以上框架后，回复「确认」或「开始决策」。\n═══════════════════════════════════════════════\n```\n\n---\n\n## T3 — 第1轮汇总模板\n\n第1轮独立评估结束后，**立即输出以下内容**：\n\n```\n═══════════════════════════════════════════════\n📊 第 1 轮独立评估 — 评分汇总\n═══════════════════════════════════════════════\n\n【评审类型】：决策点拆分评审\n【评委到齐】：{X}/{X}（{缺席评委名} 超时未提交）\n\n───────────────────────────────────────────────\n【本轮判定结果】：{通过 / 待定 / 不通过}\n阈值设置：通过≥90% / 待定[60%,90%) / 否定<60%\n═══════════════════════════════════════════════\n```\n\n### 第一轮决策项目展示\n\n**决策点拆分评审**（评委对每个决策点评分，加权平均）：\n\n| 决策点（权重） | {决策点1}（{X}%） | {决策点2}（{X}%） | … | 加权后分数 |\n|:---|:---:|:---:|:---:|:---:|\n| {模型A}（实名） | {X}/100 | {X}/100 | … | **{X}/100** |\n| {模型B}（实名） | {X}/100 | {X}/100 | … | **{X}/100** |\n| **评委均分** | {X}/100 | {X}/100 | … | **{X}/100** |\n| **通过情况** | {✅/🟡/🔴} | {✅/🟡/🔴} | … | — |\n\n───────────────────────────────────────────────\n已通过决策点已封存，不再进入下轮讨论。\n本轮🟡待定项：{决策点1}（{X}%）……由组织者综合评委建议，形成调整优化方向后，进入第二轮分歧讨论。\n本轮🔴不通过项：{决策点2}（{X}%）……由组织者综合评委建议，形成调整优化方向后，进入第二轮分歧讨论。\n请问是否「开始」或「继续」进入第二轮讨论？还是「结束」，直接结束讨论，直接输出最终结果。您可以说「开始」或「继续」，也可以说「结束」。\n\n---\n\n## T5 — 第2轮汇总模板\n\n第2轮分歧讨论结束后，**立即输出以下内容**：\n\n```\n═══════════════════════════════════════════════\n📊 第 2 轮分歧讨论 — 评分汇总\n═══════════════════════════════════════════════\n\n【本轮讨论范围】：仅第 1 轮🟡待定及🔴不通过决策点\n【评委到齐】：{X}/{X}（{缺席评委名} 超时未提交）\n\n───────────────────────────────────────────────\n【本轮判定结果】：{通过 / 待定 / 不通过}\n阈值设置：通过≥90% / 待定[60%,90%) / 否定<60%\n═══════════════════════════════════════════════\n```\n\n### 第二轮决策项目展示\n\n**不通过决策点本轮评分**：\n\n| 决策点 | 评委 | 第1轮得分 | 本轮得分 | 变化 | 是否达标 |\n|:---|:---|:---:|:---:|:---:|:---:|\n| {决策点1} | {模型A}（实名） | {X} | {X} | {+/-X} | {✅/❌} |\n| {决策点1} | {模型B}（实名） | {X} | {X} | {+/-X} | {✅/❌} |\n| {决策点2} | {模型A}（实名） | {X} | {X} | {+/-X} | {✅/❌} |\n| {决策点2} | {模型B}（实名） | {X} | {X} | {+/-X} | {✅/❌} |\n\n**评委立场变化追踪**：\n\n| 评委 | 决策点1立场变化 | 决策点2立场变化 |\n|:---|:---|:---|\n| {模型A}（实名） | {坚持/改变/部分同意} | {坚持/改变/部分同意} |\n| {模型B}（实名） | {坚持/改变/部分同意} | {坚持/改变/部分同意} |\n\n───────────────────────────────────────────────\n已通过决策点已封存，不再进入下轮讨论。\n本轮🟡待定项：{决策点1}（{X}%）……由组织者综合评委建议，形成调整优化方向后，进入第三轮（如有）或最终报告。\n本轮🔴不通过项：{决策点2}（{X}%）……由组织者综合评"}],"languages":[],"docsSourceLabel":"CLAWHUB","editorialOverview":"多模型决策委员会 — 消除单模型偏见，通过多轮分歧讨论产出客观决策参考。支持3-6个模型同时评审，提供量化投票矩阵和6段式共识报告。触发条件：包含「多模型决策」或「多模型委员会」时自动激活。 Skill: Multi Model Consensus Owner: zeekr0808-hue Summary: 多模型决策委员会 — 消除单模型偏见，通过多轮分歧讨论产出客观决策参考。支持3-6个模型同时评审，提供量化投票矩阵和6段式共识报告。触发条件：包含「多模型决策」或「多模型委员会」时自动激活。 Tags: latest:1.9.1 Version history: v1.8.0 | 2026-05-16T14:47:53.996Z | auto **Summary:** This release enforces stricter process control and standardization for the multi-model decision committee, focusing on compliance, template usage, and evaluation transparency.","editorialQuality":{"score":100,"threshold":65,"status":"ready","wordCount":1434,"uniquenessScore":45,"reasons":[]}},"media":{"evidence":{"source":"no-media","verified":false,"confidence":"low","updatedAt":"2026-10-10T18:44:47.834Z","emptyReason":"No screenshots, media assets, or demo links are available."},"primaryImageUrl":null,"mediaAssetCount":0,"assets":[],"demoUrl":null},"ownerResources":{"evidence":{"source":"unclaimed","verified":false,"confidence":"low","updatedAt":"2026-10-10T18:44:47.834Z","emptyReason":"This page has not been claimed by the agent owner."},"hasCustomPage":false,"customPageUpdatedAt":null,"customLinks":[],"structuredLinks":{"docsUrl":null,"demoUrl":null,"supportUrl":null,"pricingUrl":null,"statusUrl":null},"customPage":null},"relatedAgents":{"evidence":{"source":"protocol-neighbors","verified":false,"confidence":"medium","updatedAt":"2026-10-10T21:55:11.741Z","emptyReason":null},"items":[{"id":"8ebccd8e-3863-4187-8355-c3f14e1f9edf","entityType":"agent","canonicalPath":"/agent/iofficeai-aionui","slug":"iofficeai-aionui","name":"AionUi","description":"Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!","url":"https://github.com/iOfficeAI/AionUi","homepage":"https://www.aionui.com","source":"GITHUB_REPOS","protocols":["MCP","OPENCLAW"],"capabilities":[],"safetyScore":100,"overallRank":70,"updatedAt":"2026-10-09T19:11:12.944Z","createdAt":"2026-02-25T03:38:16.584Z","downloads":null},{"id":"b917f68a-ebff-438e-84f8-3f4b2494c0bc","entityType":"agent","canonicalPath":"/agent/activepieces-activepieces","slug":"activepieces-activepieces","name":"activepieces","description":"AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents","url":"https://github.com/activepieces/activepieces","homepage":"https://www.activepieces.com","source":"GITHUB_REPOS","protocols":["OPENCLAW"],"capabilities":[],"safetyScore":100,"overallRank":70,"updatedAt":"2026-04-15T02:22:12.426Z","createdAt":"2026-02-25T03:38:12.412Z","downloads":null},{"id":"5cb26759-3a39-483f-94cf-276a98c13bb8","entityType":"agent","canonicalPath":"/agent/cherryhq-cherry-studio","slug":"cherryhq-cherry-studio","name":"cherry-studio","description":"AI productivity studio with smart chat, autonomous agents, and 300+ assistants. Unified access to frontier LLMs","url":"https://github.com/CherryHQ/cherry-studio","homepage":"https://cherry-ai.com","source":"GITHUB_REPOS","protocols":["MCP","OPENCLAW"],"capabilities":[],"safetyScore":100,"overallRank":70,"updatedAt":"2026-04-11T14:38:40.986Z","createdAt":"2026-02-25T03:38:19.379Z","downloads":null},{"id":"6f6582d0-5d76-4f0f-b81d-86520247950b","entityType":"agent","canonicalPath":"/agent/copilotkit-copilotkit","slug":"copilotkit-copilotkit","name":"CopilotKit","description":"The Frontend for Agents & Generative UI. React + Angular","url":"https://github.com/CopilotKit/CopilotKit","homepage":"https://docs.copilotkit.ai","source":"GITHUB_REPOS","protocols":["OPENCLAW"],"capabilities":[],"safetyScore":100,"overallRank":70,"updatedAt":"2026-03-25T09:50:57.846Z","createdAt":"2026-02-25T03:39:14.617Z","downloads":null}],"links":{"hub":"/agent","source":"/agent/source/clawhub","protocols":[{"label":"OpenClaw","href":"/agent/protocol/openclew"}]}}}