{"id":"cf280e5f-45b8-41da-a528-759757ba9980","entityType":"agent","slug":"clawhub-18072937735-smyx-visual-summary-analysis","name":"Visual Summarization Skill | 视觉摘要智述技能","canonicalUrl":"https://www.xpersona.co/agent/clawhub-18072937735-smyx-visual-summary-analysis","canonicalPath":"/agent/clawhub-18072937735-smyx-visual-summary-analysis","generatedAt":"2026-10-10T02:03:04.765Z","source":"CLAWHUB","claimStatus":"UNCLAIMED","verificationTier":"NONE","summary":{"evidence":{"source":"editorial-content","verified":true,"confidence":"high","updatedAt":"2026-10-09T15:52:45.404Z","emptyReason":null},"description":"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 Skill: Visual Summarization Skill | 视觉摘要智述技能 Owner: 18072937735 Summary: Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 Tags: latest:1.0.17 Version history: v1.0.17 | 2026-10-03T22:23:26.880Z | auto - Bumped version to 1.0.18. - Updated SKILL.md with latest information and minor clarifications. - Removed redunda","descriptionLabel":"Technical summary","evidenceSummary":"Capability contract not published. No trust telemetry is available yet. 2.4K downloads reported by the source. Last updated 10/9/2026.","installCommand":"clawhub skill install s17f8q65zg3y98t86jdg1177g583whq8:smyx-visual-summary-analysis","sourceUrl":"https://clawhub.ai/18072937735/smyx-visual-summary-analysis","homepage":"https://clawhub.ai/18072937735/skills/smyx-visual-summary-analysis","primaryLinks":[{"label":"View on ClawHub","url":"https://clawhub.ai/18072937735/smyx-visual-summary-analysis","kind":"source"},{"label":"Homepage","url":"https://clawhub.ai/18072937735/skills/smyx-visual-summary-analysis","kind":"homepage"}],"safetyScore":84,"overallRank":62,"popularityScore":67,"trustScore":null,"claimedByName":null,"isOwner":false,"seoDescription":"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 Skill:"},"coverage":{"evidence":{"source":"public-profile","verified":false,"confidence":"medium","updatedAt":"2026-10-09T15:52:45.404Z","emptyReason":null},"protocols":[{"protocol":"OPENCLEW","label":"OpenClaw","status":"self-declared","notes":"Declared in the public agent profile."}],"capabilities":[],"verifiedCount":0,"selfDeclaredCount":1,"capabilityMatrix":{"rows":[{"key":"OPENCLEW","type":"protocol","support":"unknown","confidenceSource":"profile","notes":"Listed on profile"}],"flattenedTokens":"protocol:OPENCLEW|unknown|profile"}},"adoption":{"evidence":{"source":"CLAWHUB","verified":false,"confidence":"medium","updatedAt":"2026-10-09T15:52:45.404Z","emptyReason":null},"stars":null,"forks":null,"downloads":2360,"packageName":null,"latestVersion":"1.0.17","tractionLabel":"2.4K downloads"},"release":{"evidence":{"source":"CLAWHUB","verified":false,"confidence":"medium","updatedAt":"2026-10-09T15:52:45.403Z","emptyReason":null},"lastUpdatedAt":"2026-10-09T15:52:45.404Z","lastCrawledAt":"2026-10-09T15:52:45.403Z","lastIndexedAt":null,"nextCrawlAt":"2026-10-10T15:52:45.403Z","lastVerifiedAt":null,"highlights":[{"version":"1.0.17","createdAt":"2026-10-03T22:23:26.880Z","changelog":"- Bumped version to 1.0.18. - Updated SKILL.md with latest information and minor clarifications. - Removed redundant documentation file: skill-card.md.","fileCount":31,"zipByteSize":38422},{"version":"1.0.16","createdAt":"2026-09-23T18:43:51.037Z","changelog":"Version 1.0.17 - Updated SKILL.md: bumped version to 1.0.17 and made content adjustments. - Updated configuration in skills/smyx_common/scripts/config.yaml. - Removed the file skill-card.md.","fileCount":31,"zipByteSize":38714},{"version":"1.0.15","createdAt":"2026-09-17T17:19:26.541Z","changelog":"- Bumped the version to 1.0.16. - Updated documentation in SKILL.md; various clarifications, minor corrections, and updates. - Removed the file skill-card.md.","fileCount":31,"zipByteSize":38701},{"version":"1.0.14","createdAt":"2026-08-28T02:39:33.747Z","changelog":"- Version bumped to 1.0.14. - Documentation in SKILL.md updated; content and structure improved. - Obsolete skill-card.md file removed.","fileCount":31,"zipByteSize":38746},{"version":"1.0.13","createdAt":"2026-08-23T08:43:28.336Z","changelog":"- Version updated to 1.0.13. - Updated documentation and workflow in SKILL.md. - Configuration adjustments in skills/smyx_common/scripts/config.yaml. - Removed obsolete file skill-card.md.","fileCount":31,"zipByteSize":38700},{"version":"1.0.12","createdAt":"2026-08-09T21:32:34.926Z","changelog":"- Updated version number in SKILL.md. - Removed redundant skill-card.md file. - No changes to core functionality or usage.","fileCount":31,"zipByteSize":38747},{"version":"1.0.11","createdAt":"2026-08-03T13:24:36.006Z","changelog":"- Removed the file skill-card.md. - No other changes to features, functionality, or documentation.","fileCount":31,"zipByteSize":38538},{"version":"1.0.10","createdAt":"2026-08-02T01:20:33.729Z","changelog":"Version 1.0.10 - Updated SKILL.md with content or documentation changes. - Removed obsolete skill-card.md file. - No changes to core functionality or interface.","fileCount":31,"zipByteSize":38407}]},"execution":{"evidence":{"source":"CLAWHUB","verified":false,"confidence":"low","updatedAt":null,"emptyReason":"No published capability contract is available yet."},"installCommand":"clawhub skill install s17f8q65zg3y98t86jdg1177g583whq8:smyx-visual-summary-analysis","setupComplexity":"low","setupSteps":["Setup complexity is classified as HIGH. You must provision dedicated cloud infrastructure or an isolated VM. Do not run this directly on your local workstation.","Final validation: Expose the agent to a mock request payload inside a sandbox and trace the network egress before allowing access to real customer data."],"contract":{"contractStatus":"missing","authModes":[],"requires":[],"forbidden":[],"supportsMcp":false,"supportsA2a":false,"supportsStreaming":false,"inputSchemaRef":null,"outputSchemaRef":null,"dataRegion":null,"contractUpdatedAt":null,"sourceUpdatedAt":null,"freshnessSeconds":null},"invocationGuide":{"preferredApi":{"snapshotUrl":"https://www.xpersona.co/api/v1/agents/clawhub-18072937735-smyx-visual-summary-analysis/snapshot","contractUrl":"https://www.xpersona.co/api/v1/agents/clawhub-18072937735-smyx-visual-summary-analysis/contract","trustUrl":"https://www.xpersona.co/api/v1/agents/clawhub-18072937735-smyx-visual-summary-analysis/trust"},"curlExamples":["curl -s \"https://www.xpersona.co/api/v1/agents/clawhub-18072937735-smyx-visual-summary-analysis/snapshot\"","curl -s \"https://www.xpersona.co/api/v1/agents/clawhub-18072937735-smyx-visual-summary-analysis/contract\"","curl -s \"https://www.xpersona.co/api/v1/agents/clawhub-18072937735-smyx-visual-summary-analysis/trust\""],"jsonRequestTemplate":{"query":"summarize this repo","constraints":{"maxLatencyMs":2000,"protocolPreference":["OPENCLEW"]}},"jsonResponseTemplate":{"ok":true,"result":{"summary":"...","confidence":0.9},"meta":{"source":"CLAWHUB","generatedAt":"2026-10-10T02:03:04.762Z"}},"retryPolicy":{"maxAttempts":3,"backoffMs":[500,1500,3500],"retryableConditions":["HTTP_429","HTTP_503","NETWORK_TIMEOUT"]}},"endpoints":{"dossierUrl":"https://www.xpersona.co/api/v1/agents/clawhub-18072937735-smyx-visual-summary-analysis/dossier","snapshotUrl":"https://www.xpersona.co/api/v1/agents/clawhub-18072937735-smyx-visual-summary-analysis/snapshot","contractUrl":"https://www.xpersona.co/api/v1/agents/clawhub-18072937735-smyx-visual-summary-analysis/contract","trustUrl":"https://www.xpersona.co/api/v1/agents/clawhub-18072937735-smyx-visual-summary-analysis/trust"}},"reliability":{"evidence":{"source":"runtime-metrics","verified":false,"confidence":"low","updatedAt":null,"emptyReason":"No trust, reliability, or runtime telemetry is available."},"trust":{"status":"unavailable","handshakeStatus":"UNKNOWN","verificationFreshnessHours":null,"reputationScore":null,"p95LatencyMs":null,"successRate30d":null,"fallbackRate":null,"attempts30d":null,"trustUpdatedAt":null,"trustConfidence":"unknown","sourceUpdatedAt":null,"freshnessSeconds":null},"decisionGuardrails":{"doNotUseIf":["Contract metadata is missing or unavailable for deterministic execution."],"safeUseWhen":[],"riskFlags":["missing_or_unavailable_contract","trust_data_unavailable","schema_references_missing"],"operationalConfidence":"low"},"executionMetrics":{"observedLatencyMsP50":null,"observedLatencyMsP95":null,"estimatedCostUsd":null,"uptime30d":null,"rateLimitRpm":null,"rateLimitBurst":null,"lastVerifiedAt":null,"verificationSource":null},"runtimeMetrics":{"successRate":null,"avgLatencyMs":null,"avgCostUsd":null,"hallucinationRate":null,"retryRate":null,"disputeRate":null,"p50Latency":null,"p95Latency":null,"lastUpdated":null}},"benchmarks":{"evidence":{"source":"no-benchmark-data","verified":false,"confidence":"low","updatedAt":null,"emptyReason":"No benchmark suites or observed failure patterns are available."},"suites":[],"failurePatterns":[]},"artifacts":{"evidence":{"source":"CLAWHUB","verified":false,"confidence":"high","updatedAt":"2026-10-09T15:52:45.404Z","emptyReason":null},"readme":"Skill: Visual Summarization Skill | 视觉摘要智述技能\n\nOwner: 18072937735\n\nSummary: Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\n\nTags: latest:1.0.17\n\nVersion history:\n\nv1.0.17 | 2026-10-03T22:23:26.880Z | auto\n\n- Bumped version to 1.0.18.\n- Updated SKILL.md with latest information and minor clarifications.\n- Removed redundant documentation file: skill-card.md.\n\nv1.0.16 | 2026-09-23T18:43:51.037Z | auto\n\nVersion 1.0.17\n\n- Updated SKILL.md: bumped version to 1.0.17 and made content adjustments.\n- Updated configuration in skills/smyx_common/scripts/config.yaml.\n- Removed the file skill-card.md.\n\nv1.0.15 | 2026-09-17T17:19:26.541Z | auto\n\n- Bumped the version to 1.0.16.\n- Updated documentation in SKILL.md; various clarifications, minor corrections, and updates.\n- Removed the file skill-card.md.\n\nv1.0.14 | 2026-08-28T02:39:33.747Z | auto\n\n- Version bumped to 1.0.14.\n- Documentation in SKILL.md updated; content and structure improved.\n- Obsolete skill-card.md file removed.\n\nv1.0.13 | 2026-08-23T08:43:28.336Z | auto\n\n- Version updated to 1.0.13.\n- Updated documentation and workflow in SKILL.md.\n- Configuration adjustments in skills/smyx_common/scripts/config.yaml.\n- Removed obsolete file skill-card.md.\n\nv1.0.12 | 2026-08-09T21:32:34.926Z | auto\n\n- Updated version number in SKILL.md.\n- Removed redundant skill-card.md file.\n- No changes to core functionality or usage.\n\nv1.0.11 | 2026-08-03T13:24:36.006Z | auto\n\n- Removed the file skill-card.md.\n- No other changes to features, functionality, or documentation.\n\nv1.0.10 | 2026-08-02T01:20:33.729Z | auto\n\nVersion 1.0.10\n\n- Updated SKILL.md with content or documentation changes.\n- Removed obsolete skill-card.md file.\n- No changes to core functionality or interface.\n\nv1.0.9 | 2026-07-19T05:08:31.165Z | auto\n\n- Bumped version from 1.0.7 to 1.0.8 in SKILL.md.\n- Removed the file skill-card.md.\n- Made updates in skills/smyx_common/scripts/util.py (details not shown).\n- Minor documentation and maintenance updates.\n\nv1.0.8 | 2026-07-08T11:49:27.819Z | auto\n\n- Updated SKILL.md to version 1.0.7 with minor documentation adjustments.\n- Removed redundant or unnecessary skill-card.md file.\n- No changes to core functionality or features; documentation only.\n\nv1.0.7 | 2026-07-07T18:48:53.536Z | auto\n\n- Updated version to 1.0.6 in SKILL.md.\n- Minor revisions in documentation for accuracy and clarity.\n- Removed the obsolete skill-card.md file.\n\nv1.0.6 | 2026-06-30T12:20:26.713Z | auto\n\n- Internal code improvements in dao.py and util.py for better structure or efficiency.\n- Documentation cleanup: Removed the redundant skill-card.md file.\n- SKILL.md: No user-facing or functional changes in skill documentation/content.\n\nv1.0.5 | 2026-06-24T20:28:26.831Z | auto\n\n- Major update: Skill usage flow and user identity handling redesigned for a smoother experience.\n- User identity (open-id) is now handled automatically and securely by the backend; users are never asked for or shown identity parameters.\n- Updated documentation to emphasize that all history/report queries must retrieve data from the cloud API; local or manual lookups are strictly prohibited.\n- Changelog, usage instructions, and workflow clarified and streamlined with clear tables and visuals.\n- Removed unnecessary files and deprecated open-id guidance from user-facing materials.\n\nv1.0.4 | 2026-05-25T09:04:42.103Z | auto\n\n- Separated the \"smyx_analysis\" skill logic and code into its own module and directory, improving code organization.\n- Removed the previous \"face_analysis\" skill and related files.\n- Updated and refactored core scripts and requirements to align with the modular structure.\n- Adjusted attachment and file handling for user uploads, including support for new directory structure.\n- Updated configuration, API documentation, and references to match current module layout and logic.\n- Reduced supported input file size for analysis from 100MB to 10MB.\n\nv1.0.3 | 2026-05-18T11:33:43.654Z | auto\n\n- Updated environment configuration files: added or modified config-dev.yaml, config-test.yaml, and config.yaml for improved environment management.\n- Refactored config.py to enhance configuration handling across different development stages.\n- Made utility updates in util.py for better script support.\n- No changes to the skill’s user-facing behavior or workflow.\n- Documentation content in SKILL.md remains unchanged.\n\nv1.0.2 | 2026-05-12T06:23:38.635Z | auto\n\nVersion 1.0.2\n\n- Updated configuration files: config.yaml, config-dev.yaml, and config-test.yaml.\n- Made enhancements and refactoring in core scripts: __init__.py, config.py, and util.py.\n- No user-facing feature or workflow changes documented.\n- Documentation (SKILL.md) remains unchanged.\n\nv1.0.1 | 2026-05-07T11:13:39.801Z | auto\n\n- Added version field (\"1.0.0\") to SKILL.md for clearer versioning information.\n- No functional or logic changes to the skill's processes or usage.\n- Documentation and usage instructions remain consistent with the previous version.\n\nv1.0.0 | 2026-04-18T09:31:54.986Z | auto\n\nVisual Summarization Skill v1.0.0 – First stable release\n\n- Introduces an AI-powered skill for generating coherent scene descriptions from video clips or image content.\n- Enforces strict rules: prohibits use of local memory files and mandates all history queries from a cloud API with open-id authentication.\n- Requires users to provide a valid open-id via config file or user input before analysis.\n- Automatically extracts and summarizes scene elements such as objects, behaviors, and environments into fluent Chinese text.\n- Adds support for listing and viewing reports in a Markdown table linked to cloud report images.\n- Provides clear guidance and requirements for input quality and workflow usage.\n\nArchive index:\n\nArchive v1.0.17: 31 files, 38422 bytes\n\nFiles: references/api_doc.md (671b), scripts/__init__.py (31b), scripts/config.py (654b), scripts/config.yaml (3b), scripts/skill.py (573b), scripts/visual_summary_analysis.py (3210b), skill-card.md (2026b), SKILL.md (9793b), skills/smyx_analysis/__init__.py (0b), skills/smyx_analysis/references/api_doc.md (427b), skills/smyx_analysis/requirements.txt (45b), skills/smyx_analysis/scripts/__init__.py (0b), skills/smyx_analysis/scripts/api_service.py (1509b), skills/smyx_analysis/scripts/config.py (1003b), skills/smyx_analysis/scripts/config.yaml (3b), skills/smyx_analysis/scripts/skill.py (6529b), skills/smyx_analysis/scripts/smyx_analysis.py (3833b), skills/smyx_common/__init__.py (0b), skills/smyx_common/requirements.txt (47b), skills/smyx_common/scripts/__init__.py (177b), skills/smyx_common/scripts/api_service.py (2645b), skills/smyx_common/scripts/base.py (469b), skills/smyx_common/scripts/config-dev.yaml (214b), skills/smyx_common/scripts/config-prod.yaml (0b), skills/smyx_common/scripts/config-test.yaml (256b), skills/smyx_common/scripts/config.py (24363b), skills/smyx_common/scripts/config.yaml (473b), skills/smyx_common/scripts/dao.py (18266b), skills/smyx_common/scripts/skill.py (2473b), skills/smyx_common/scripts/util.py (28776b), _meta.json (148b)\n\nFile v1.0.17:SKILL.md\n\n---\nname: \"visual-summary-analysis\"\ndescription: \"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\"\nversion: \"1.0.18\"\nlicense: \"MIT-0\"\n---\n\n# 📝 Visual Summarization Skill | 视觉摘要智述技能\n> **智能分析中枢** · 图片/视频智能分析 · 结构化报告 · 历史报告云端查询\n\n---\n\n## 🧭 技能概览 | Overview\n\n| 模块 | 内容 |\n|---|---|\n| 🏷️ 技能名称 | **视觉摘要智述技能** |\n| 🎯 核心目标 | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 |\n| 🖼️ 输入类型 | 图片、视频、本地文件、网络 URL |\n| 📝 输出能力 | 结构化分析报告、识别/监测结果、建议与报告链接 |\n| 🧩 场景码 | `VISUAL_SUMMARY` |\n\nBased on advanced multimodal large models and video understanding technologies, this feature performs deep semantic\nanalysis and logical reasoning on input video clips or images. Utilizing computer vision algorithms, the system\nprecisely identifies key visual elements—including subject objects, environmental backgrounds, action behaviors, and\nlighting atmosphere. It then combines this with Natural Language Generation (NLG) technology to transform abstract\nvisual information into smooth, logically coherent scene descriptions. Whether dealing with dynamic video events or\nstatic image moments, the system captures critical details and restores the on-site context with vivid language. This\nprovides intelligent text summarization services for scenarios such as video content understanding, accessibility\nassistance, and media asset management.\n\n本功能基于先进的多模态大模型与视频理解技术，能够对传入的视频片段或图片进行深度语义分析与逻辑推理。系统通过计算机视觉算法精准识别画面中的主体对象、环境背景、动作行为及光影氛围，并结合自然语言生成技术，将抽象的视觉信息转化为一段通顺自然、逻辑连贯的场景描述。无论是动态的视频事件还是静态的图像瞬间，系统都能捕捉关键细节，用生动的语言还原现场情境，为视频内容理解、无障碍辅助、媒体资产管理等场景提供智能化的文本摘要服务\n\n## 🎬 技能演示 | Skill Demo\n\n[▶️ 点击查看技能使用介绍](https://lifeemergence.com/sample.html)\n\n---\n\n## 🎯 任务目标 | Goals\n\n### 1. 🧩 技能用途\n\n对传入的视频片段或图片内容进行AI分析，自动生成通顺自然的场景描述摘要\n\n### 2. 🛠️ 能力范围\n\n| 序号 | 具体能力 |\n|---:|---|\n| 1 | 场景内容识别 |\n| 2 | 物体识别 |\n| 3 | 行为识别 |\n| 4 | 文字提取 |\n| 5 | 整合成一段流畅自然的中文描述 |\n\n### 3. ⚡ 触发条件\n\n| 触发类型 | 触发规则 |\n|---|---|\n| ✅ 默认触发 | **默认触发**：当用户提供视频/图片需要生成内容描述/视觉摘要时，默认触发本技能 |\n| 🔎 明确分析意图 | 当用户明确需要视频内容描述、图片内容摘要、视觉智述时，提及视频摘要、内容描述、视觉摘要智述、视频转文字等关键词，并且上传了视频/图片 |\n| 📚 历史报告查询 | 当用户提及以下关键词时，**自动触发历史报告查询功能** ：查看历史摘要报告、摘要报告清单、报告列表、查询历史摘要报告、显示所有摘要报告、视觉智述分析报告，查询视觉摘要智述分析报告 |\n\n### 4. 🤖 自动行为\n\n| 自动行为 | 执行要求 |\n|---|---|\n| 📎 附件处理 | 如果用户上传了附件或者视频/图片文件，则自动保存为本地文件 |\n| ☁️ 历史报告查询 | 如果用户触发历史报告查询关键词，必须直接调用云端 API 查询，不得从本地记忆或人工汇总中获取 |\n\n#### ⚠️ 强制数据获取规则（次高优先级）\n\n> **橙色强约束：** 历史报告清单只允许从云端接口读取，不允许从本地记录、长期记忆或人工汇总中提取。\n\n必须执行：\n\n```bash\npython -m scripts.visual_summary_analysis --list\n```\n\n| 类型 | 要求 |\n|---|---|\n| ✅ 必须 | 使用 `python -m scripts.visual_summary_analysis --list` 调用 API 查询云端的历史报告数据 |\n| 🚫 严格禁止 | 从本地 `memory` 目录读取历史会话信息 |\n| 🚫 严格禁止 | 手动汇总本地记录中的报告 |\n| 🚫 严格禁止 | 从长期记忆中提取报告 |\n| ✅ 输出格式 | 必须统一从云端接口获取最新完整数据，然后以 Markdown 表格格式输出结果 |\n\n## 📦 前置准备 | Requirements\n- 依赖说明:scripts 脚本所需的依赖包及版本\n  ```\n  requests>=2.28.0\n  ```\n\n## 📸 使用要求 | Usage Requirements\n| 要求项 | 说明 |\n|---|---|\n| 视频/图片内容清晰 | ，主要物体和场景完整可见 |\n| 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 | 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 |\n| 主要场景主体不被大面积遮挡 | 主要场景主体不被大面积遮挡 |\n\n## 🚀 操作步骤 | Workflow\n### 🔐 用户身份处理（内部自动完成）\n\n> **绿色安全原则：** 用户身份参数由系统内部自动处理，**不得向用户展示、询问或要求输入任何身份标识**。\n\n执行本技能分析或历史报告查询时，脚本会自动完成身份初始化：\n\n| 场景 | 系统行为 |\n|---|---|\n| 上游系统有内部身份参数 | 由脚本静默接收并使用 |\n| 上游系统未提供内部身份参数 | 脚本会自动复用本地缺省用户 |\n| 本地缺省用户不存在 | 脚本会自动创建并在后续任务中复用 |\n| 对用户输出 | 只展示分析进度、分析结果和报告链接，不展示内部身份值 |\n\n#### 🔒 关键约束\n\n| 禁止/要求 | 说明 |\n|---|---|\n| 🚫 不得询问身份 | 不得提示用户输入用户名、手机号或任何内部身份参数 |\n| 🚫 不得暴露身份值 | 不得在回复、报告、示例、错误提示中暴露内部身份值 |\n| 🚫 不得列为用户参数 | 不得把内部身份参数列为用户需要理解或传入的参数 |\n| ✅ 自动关联报告 | 历史报告查询同样由系统内部身份自动关联，用户只需表达“查看历史报告/报告清单”等意图 |\n\n---\n\n### 🧪 标准流程 | Standard Flow\n\n| 步骤 | 阶段 | 执行动作 |\n|---:|---|---|\n| 1 | 📥 准备视频/图片输入 | 提供本地文件路径或网络 URL；确保输入内容清晰、符合技能场景要求 |\n| 2 | 🔐 系统自动完成身份关联 | 无需用户输入任何身份参数；不在回复中展示内部身份值 |\n| 3 | ⚙️ 执行视觉摘要智述分析 | 调用 `-m scripts.visual_summary_analysis` 处理输入（**必须在技能根目录下运行脚本**） |\n| 4 | 📊 查看分析结果 | 接收结构化分析报告，查看识别/监测结果、风险提示、建议与报告链接 |\n\n### ⚙️ 脚本参数说明\n\n| 参数 | 含义 | 备注 |\n|---|---|---|\n| `--input` | 本地视频/图片文件路径 | 适用于本地文件分析 |\n| `--url` | 网络视频/图片 URL 地址（API 服务自动下载） | API 服务自动下载网络资源 |\n| `--list` | 显示历史视觉摘要智述分析报告列表清单（可以输入起始日期参数过滤数据范围） | 用于云端历史报告查询 |\n| `--api-url` | API 服务地址（可选，使用默认值） | 按需填写 |\n| `--detail` | 输出详细程度（basic/standard/json，默认 json） | 输出详细程度 |\n| `--output` | 结果输出文件路径（可选） | 可选 |\n\n## 🗂️ 资源索引 | Resource Index\n| 资源类型 | 路径 | 用途 | 何时读取 |\n|---|---|---|---|\n| 🐍 必要脚本 | [`scripts/visual_summary_analysis.py`](scripts/visual_summary_analysis.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 🐍 必要脚本 | [`scripts/config.py`](scripts/config.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 📘 领域参考 | [`references/api_doc.md`](references/api_doc.md) | 了解 API 接口规范、字段说明和错误码 | 仅在需要了解接口规范或错误码时读取 |\n\n## ⚠️ 注意事项 | Notes\n| 分类 | 注意事项 |\n|---|---|\n| 📚 文档读取 | 仅在需要时读取参考文档，保持上下文简洁 |\n| 📁 格式支持 | 支持格式：jpg/jpeg/png/mp4/avi/mov，最大 10MB |\n| 🚫 脚本限制 | 禁止临时生成脚本，只能用技能本身的脚本 |\n| 🌐 网络地址 | 传入的网路地址参数，不需要下载本地，默认地址都是公网地址，api 服务会自动下载 |\n| 📜 报告输出 | 当显示历史分析报告清单的时候，从接口返回 json 数据中提取字段  作为超链接地址，且自动转化为如下 Markdown |\n| 📜 报告输出 | 表格输出示例 |\n\n## 🧰 使用示例 | Examples\n```bash\n# 分析本地视频片段\npython -m scripts.visual_summary_analysis --input /path/to/clip.mp4 分析本地图片\npython -m scripts.visual_summary_analysis --input /path/to/image.jpg 分析网络视频\npython -m scripts.visual_summary_analysis --url https://example.com/clip.mp4 显示历史摘要报告/显示摘要报告清单列表/显示历史智述（自动触发关键词：查看历史摘要报告、历史报告、摘要报告清单等）\npython -m scripts.visual_summary_analysis --list\n\n# 输出精简报告\npython -m scripts.visual_summary_analysis --input clip.mp4 --detail basic\n\n# 保存结果到文件\npython -m scripts.visual_summary_analysis --input clip.mp4 --output result.json\n```\n\nFile v1.0.17:_meta.json\n\n{\n  \"ownerId\": \"kn7e2caqj7pnsvr9r7t8zenghs83xw7n\",\n  \"slug\": \"smyx-visual-summary-analysis\",\n  \"version\": \"1.0.17\",\n  \"publishedAt\": 1791066206880\n}\n\nFile v1.0.17:references/api_doc.md\n\n# API 接口文档\n\n此处用于存放视觉摘要智述分析 API 的接口文档，待后续补充。\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 主要接口\n\n1. `/web/ai-analysis/v2/start-common-ai-analysis` - 启动AI分析任务\n2. `/web/ai-analysis/v2/get-common-ai-analysis-result` - 获取分析结果\n3. `/web/ai-analysis/page-common-ai-analysis-result` - 分页查询历史报告\n4. `/ai/order/api/getReportDetailExport?id={id}` - 导出完整报告\n\n## 场景代码\n\n- `OPEN_VISUAL_SUMMARY_ANALYSIS` - 开放平台视觉摘要智述分析\n\nFile v1.0.17:skills/smyx_analysis/references/api_doc.md\n\n# API接口文档\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 错误码说明\n\n| 错误码 | 说明       |\n|-----|----------|\n| 400 | 请求参数错误   |\n| 401 | API密钥无效  |\n| 403 | 权限不足     |\n| 413 | 文件过大     |\n| 415 | 不支持的文件格式 |\n| 500 | 服务器内部错误  |\n\nFile v1.0.17:scripts/config.yaml\n\n{}\n\nFile v1.0.17:skills/smyx_analysis/scripts/config.yaml\n\n{}\n\nFile v1.0.17:skills/smyx_common/scripts/config-dev.yaml\n\nApiEnum:\n  base-url-open-api: \"http://192.168.1.234:9601/smyx-open-api\"\n  base-url-open-h5: \"http://192.168.1.234:4100\"\n  base-url-health: \"http://192.168.1.234:7070/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.17:skills/smyx_common/scripts/config-test.yaml\n\nApiEnum:\n  base-url-open-api: \"https://livemonitortest.lifeemergence.com/smyx-open-api\"\n  base-url-open-h5: \"http://livemonitortest.lifeemergence.com\"\n  base-url-health: \"https://healthtest.lifeemergence.com/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.17:skills/smyx_common/scripts/config.yaml\n\nApiEnum:\n  api-key: null\n  api-secret-key: null\n  base-url-health: https://lifeemergence.com/jeecg-boot-xzgz\n  base-url-open-api: https://open.lifeemergence.com/smyx-open-api\n  base-url-open-h5: http://livemonitor.lifeemergence.com\n  database-url: null\nConstantEnum:\n  app--id: x1a3s4nwy1s2r4se\n  current--tentant-code: XIAN_ZHAO_GAN_ZHI\n  default--skill-platform-name: ARK_CLAW\n  feishu-app--id: cli_a93d769369badcb1\n  feishu-app--secret: null\n  is-debug: false\nenv: prod\n\nFile v1.0.17:skill-card.md\n\n## Description:\n\nAnalyzes images and video to produce natural-language scene descriptions and structured reports.\n\nThis skill is ready for commercial/non-commercial use.\n\n## Publisher:\n\n[18072937735](https://clawhub.ai/user/18072937735)\n\n### License/Terms of Use:\n\nMIT-0\n\n## Use Case:\n\nDevelopers and media teams use this skill to summarize images or video clips, review scene analysis, and retrieve prior reports.\n\n### Deployment Geography for Use:\n\nGlobal\n\n## Known Risks and Mitigations:\n\nRisk: Images, videos, or their URLs are sent to a remote service and may contain sensitive information.\n\nMitigation: Share only media authorized for remote processing; review data-handling requirements before use.\n\nRisk: The skill silently creates or reuses a cloud identity and can retrieve reports linked to it.\n\nMitigation: Obtain user approval for cloud account linkage and report-history access before enabling the skill.\n\nRisk: Authentication tokens persist in the local workspace data directory.\n\nMitigation: Protect access to that directory and remove or rotate stored tokens when access is no longer needed.\n\n## Reference(s):\n\n- [ClawHub skill release](https://clawhub.ai/18072937735/skills/smyx-visual-summary-analysis)\n- [Visual summary API documentation](artifact/references/api_doc.md)\n- [Skill demonstration](https://lifeemergence.com/sample.html)\n\n## Skill Output:\n\n**Output Type(s):** [Text, Markdown, Files]\n\n**Output Format:** [Natural-language scene description or structured report; report history as a Markdown table]\n\n**Output Parameters:** [1D]\n\n**Other Properties Related to Output:** [May include report links and optionally save the result to a file.]\n\n## Skill Version(s):\n\n1.0.17 (source: ClawHub release; bundled frontmatter says 1.0.18)\n\n## Ethical Considerations:\n\nUsers should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.\n\nFile v1.0.17:skills/smyx_analysis/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nyaml==6.0.3\n\nFile v1.0.17:skills/smyx_common/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nPyYAML==6.0.3\n\nArchive v1.0.16: 31 files, 38714 bytes\n\nFiles: references/api_doc.md (671b), scripts/__init__.py (31b), scripts/config.py (654b), scripts/config.yaml (3b), scripts/skill.py (573b), scripts/visual_summary_analysis.py (3210b), skill-card.md (2445b), SKILL.md (9793b), skills/smyx_analysis/__init__.py (0b), skills/smyx_analysis/references/api_doc.md (427b), skills/smyx_analysis/requirements.txt (45b), skills/smyx_analysis/scripts/__init__.py (0b), skills/smyx_analysis/scripts/api_service.py (1509b), skills/smyx_analysis/scripts/config.py (1003b), skills/smyx_analysis/scripts/config.yaml (3b), skills/smyx_analysis/scripts/skill.py (6529b), skills/smyx_analysis/scripts/smyx_analysis.py (3833b), skills/smyx_common/__init__.py (0b), skills/smyx_common/requirements.txt (47b), skills/smyx_common/scripts/__init__.py (177b), skills/smyx_common/scripts/api_service.py (2645b), skills/smyx_common/scripts/base.py (469b), skills/smyx_common/scripts/config-dev.yaml (214b), skills/smyx_common/scripts/config-prod.yaml (0b), skills/smyx_common/scripts/config-test.yaml (256b), skills/smyx_common/scripts/config.py (24363b), skills/smyx_common/scripts/config.yaml (473b), skills/smyx_common/scripts/dao.py (18266b), skills/smyx_common/scripts/skill.py (2473b), skills/smyx_common/scripts/util.py (28776b), _meta.json (148b)\n\nFile v1.0.16:SKILL.md\n\n---\nname: \"visual-summary-analysis\"\ndescription: \"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\"\nversion: \"1.0.17\"\nlicense: \"MIT-0\"\n---\n\n# 📝 Visual Summarization Skill | 视觉摘要智述技能\n> **智能分析中枢** · 图片/视频智能分析 · 结构化报告 · 历史报告云端查询\n\n---\n\n## 🧭 技能概览 | Overview\n\n| 模块 | 内容 |\n|---|---|\n| 🏷️ 技能名称 | **视觉摘要智述技能** |\n| 🎯 核心目标 | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 |\n| 🖼️ 输入类型 | 图片、视频、本地文件、网络 URL |\n| 📝 输出能力 | 结构化分析报告、识别/监测结果、建议与报告链接 |\n| 🧩 场景码 | `VISUAL_SUMMARY` |\n\nBased on advanced multimodal large models and video understanding technologies, this feature performs deep semantic\nanalysis and logical reasoning on input video clips or images. Utilizing computer vision algorithms, the system\nprecisely identifies key visual elements—including subject objects, environmental backgrounds, action behaviors, and\nlighting atmosphere. It then combines this with Natural Language Generation (NLG) technology to transform abstract\nvisual information into smooth, logically coherent scene descriptions. Whether dealing with dynamic video events or\nstatic image moments, the system captures critical details and restores the on-site context with vivid language. This\nprovides intelligent text summarization services for scenarios such as video content understanding, accessibility\nassistance, and media asset management.\n\n本功能基于先进的多模态大模型与视频理解技术，能够对传入的视频片段或图片进行深度语义分析与逻辑推理。系统通过计算机视觉算法精准识别画面中的主体对象、环境背景、动作行为及光影氛围，并结合自然语言生成技术，将抽象的视觉信息转化为一段通顺自然、逻辑连贯的场景描述。无论是动态的视频事件还是静态的图像瞬间，系统都能捕捉关键细节，用生动的语言还原现场情境，为视频内容理解、无障碍辅助、媒体资产管理等场景提供智能化的文本摘要服务\n\n## 🎬 技能演示 | Skill Demo\n\n[▶️ 点击查看技能使用介绍](https://lifeemergence.com/sample.html)\n\n---\n\n## 🎯 任务目标 | Goals\n\n### 1. 🧩 技能用途\n\n对传入的视频片段或图片内容进行AI分析，自动生成通顺自然的场景描述摘要\n\n### 2. 🛠️ 能力范围\n\n| 序号 | 具体能力 |\n|---:|---|\n| 1 | 场景内容识别 |\n| 2 | 物体识别 |\n| 3 | 行为识别 |\n| 4 | 文字提取 |\n| 5 | 整合成一段流畅自然的中文描述 |\n\n### 3. ⚡ 触发条件\n\n| 触发类型 | 触发规则 |\n|---|---|\n| ✅ 默认触发 | **默认触发**：当用户提供视频/图片需要生成内容描述/视觉摘要时，默认触发本技能 |\n| 🔎 明确分析意图 | 当用户明确需要视频内容描述、图片内容摘要、视觉智述时，提及视频摘要、内容描述、视觉摘要智述、视频转文字等关键词，并且上传了视频/图片 |\n| 📚 历史报告查询 | 当用户提及以下关键词时，**自动触发历史报告查询功能** ：查看历史摘要报告、摘要报告清单、报告列表、查询历史摘要报告、显示所有摘要报告、视觉智述分析报告，查询视觉摘要智述分析报告 |\n\n### 4. 🤖 自动行为\n\n| 自动行为 | 执行要求 |\n|---|---|\n| 📎 附件处理 | 如果用户上传了附件或者视频/图片文件，则自动保存为本地文件 |\n| ☁️ 历史报告查询 | 如果用户触发历史报告查询关键词，必须直接调用云端 API 查询，不得从本地记忆或人工汇总中获取 |\n\n#### ⚠️ 强制数据获取规则（次高优先级）\n\n> **橙色强约束：** 历史报告清单只允许从云端接口读取，不允许从本地记录、长期记忆或人工汇总中提取。\n\n必须执行：\n\n```bash\npython -m scripts.visual_summary_analysis --list\n```\n\n| 类型 | 要求 |\n|---|---|\n| ✅ 必须 | 使用 `python -m scripts.visual_summary_analysis --list` 调用 API 查询云端的历史报告数据 |\n| 🚫 严格禁止 | 从本地 `memory` 目录读取历史会话信息 |\n| 🚫 严格禁止 | 手动汇总本地记录中的报告 |\n| 🚫 严格禁止 | 从长期记忆中提取报告 |\n| ✅ 输出格式 | 必须统一从云端接口获取最新完整数据，然后以 Markdown 表格格式输出结果 |\n\n## 📦 前置准备 | Requirements\n- 依赖说明:scripts 脚本所需的依赖包及版本\n  ```\n  requests>=2.28.0\n  ```\n\n## 📸 使用要求 | Usage Requirements\n| 要求项 | 说明 |\n|---|---|\n| 视频/图片内容清晰 | ，主要物体和场景完整可见 |\n| 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 | 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 |\n| 主要场景主体不被大面积遮挡 | 主要场景主体不被大面积遮挡 |\n\n## 🚀 操作步骤 | Workflow\n### 🔐 用户身份处理（内部自动完成）\n\n> **绿色安全原则：** 用户身份参数由系统内部自动处理，**不得向用户展示、询问或要求输入任何身份标识**。\n\n执行本技能分析或历史报告查询时，脚本会自动完成身份初始化：\n\n| 场景 | 系统行为 |\n|---|---|\n| 上游系统有内部身份参数 | 由脚本静默接收并使用 |\n| 上游系统未提供内部身份参数 | 脚本会自动复用本地缺省用户 |\n| 本地缺省用户不存在 | 脚本会自动创建并在后续任务中复用 |\n| 对用户输出 | 只展示分析进度、分析结果和报告链接，不展示内部身份值 |\n\n#### 🔒 关键约束\n\n| 禁止/要求 | 说明 |\n|---|---|\n| 🚫 不得询问身份 | 不得提示用户输入用户名、手机号或任何内部身份参数 |\n| 🚫 不得暴露身份值 | 不得在回复、报告、示例、错误提示中暴露内部身份值 |\n| 🚫 不得列为用户参数 | 不得把内部身份参数列为用户需要理解或传入的参数 |\n| ✅ 自动关联报告 | 历史报告查询同样由系统内部身份自动关联，用户只需表达“查看历史报告/报告清单”等意图 |\n\n---\n\n### 🧪 标准流程 | Standard Flow\n\n| 步骤 | 阶段 | 执行动作 |\n|---:|---|---|\n| 1 | 📥 准备视频/图片输入 | 提供本地文件路径或网络 URL；确保输入内容清晰、符合技能场景要求 |\n| 2 | 🔐 系统自动完成身份关联 | 无需用户输入任何身份参数；不在回复中展示内部身份值 |\n| 3 | ⚙️ 执行视觉摘要智述分析 | 调用 `-m scripts.visual_summary_analysis` 处理输入（**必须在技能根目录下运行脚本**） |\n| 4 | 📊 查看分析结果 | 接收结构化分析报告，查看识别/监测结果、风险提示、建议与报告链接 |\n\n### ⚙️ 脚本参数说明\n\n| 参数 | 含义 | 备注 |\n|---|---|---|\n| `--input` | 本地视频/图片文件路径 | 适用于本地文件分析 |\n| `--url` | 网络视频/图片 URL 地址（API 服务自动下载） | API 服务自动下载网络资源 |\n| `--list` | 显示历史视觉摘要智述分析报告列表清单（可以输入起始日期参数过滤数据范围） | 用于云端历史报告查询 |\n| `--api-url` | API 服务地址（可选，使用默认值） | 按需填写 |\n| `--detail` | 输出详细程度（basic/standard/json，默认 json） | 输出详细程度 |\n| `--output` | 结果输出文件路径（可选） | 可选 |\n\n## 🗂️ 资源索引 | Resource Index\n| 资源类型 | 路径 | 用途 | 何时读取 |\n|---|---|---|---|\n| 🐍 必要脚本 | [`scripts/visual_summary_analysis.py`](scripts/visual_summary_analysis.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 🐍 必要脚本 | [`scripts/config.py`](scripts/config.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 📘 领域参考 | [`references/api_doc.md`](references/api_doc.md) | 了解 API 接口规范、字段说明和错误码 | 仅在需要了解接口规范或错误码时读取 |\n\n## ⚠️ 注意事项 | Notes\n| 分类 | 注意事项 |\n|---|---|\n| 📚 文档读取 | 仅在需要时读取参考文档，保持上下文简洁 |\n| 📁 格式支持 | 支持格式：jpg/jpeg/png/mp4/avi/mov，最大 10MB |\n| 🚫 脚本限制 | 禁止临时生成脚本，只能用技能本身的脚本 |\n| 🌐 网络地址 | 传入的网路地址参数，不需要下载本地，默认地址都是公网地址，api 服务会自动下载 |\n| 📜 报告输出 | 当显示历史分析报告清单的时候，从接口返回 json 数据中提取字段  作为超链接地址，且自动转化为如下 Markdown |\n| 📜 报告输出 | 表格输出示例 |\n\n## 🧰 使用示例 | Examples\n```bash\n# 分析本地视频片段\npython -m scripts.visual_summary_analysis --input /path/to/clip.mp4 分析本地图片\npython -m scripts.visual_summary_analysis --input /path/to/image.jpg 分析网络视频\npython -m scripts.visual_summary_analysis --url https://example.com/clip.mp4 显示历史摘要报告/显示摘要报告清单列表/显示历史智述（自动触发关键词：查看历史摘要报告、历史报告、摘要报告清单等）\npython -m scripts.visual_summary_analysis --list\n\n# 输出精简报告\npython -m scripts.visual_summary_analysis --input clip.mp4 --detail basic\n\n# 保存结果到文件\npython -m scripts.visual_summary_analysis --input clip.mp4 --output result.json\n```\n\nFile v1.0.16:_meta.json\n\n{\n  \"ownerId\": \"kn7e2caqj7pnsvr9r7t8zenghs83xw7n\",\n  \"slug\": \"smyx-visual-summary-analysis\",\n  \"version\": \"1.0.16\",\n  \"publishedAt\": 1790189031037\n}\n\nFile v1.0.16:references/api_doc.md\n\n# API 接口文档\n\n此处用于存放视觉摘要智述分析 API 的接口文档，待后续补充。\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 主要接口\n\n1. `/web/ai-analysis/v2/start-common-ai-analysis` - 启动AI分析任务\n2. `/web/ai-analysis/v2/get-common-ai-analysis-result` - 获取分析结果\n3. `/web/ai-analysis/page-common-ai-analysis-result` - 分页查询历史报告\n4. `/ai/order/api/getReportDetailExport?id={id}` - 导出完整报告\n\n## 场景代码\n\n- `OPEN_VISUAL_SUMMARY_ANALYSIS` - 开放平台视觉摘要智述分析\n\nFile v1.0.16:skills/smyx_analysis/references/api_doc.md\n\n# API接口文档\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 错误码说明\n\n| 错误码 | 说明       |\n|-----|----------|\n| 400 | 请求参数错误   |\n| 401 | API密钥无效  |\n| 403 | 权限不足     |\n| 413 | 文件过大     |\n| 415 | 不支持的文件格式 |\n| 500 | 服务器内部错误  |\n\nFile v1.0.16:scripts/config.yaml\n\n{}\n\nFile v1.0.16:skills/smyx_analysis/scripts/config.yaml\n\n{}\n\nFile v1.0.16:skills/smyx_common/scripts/config-dev.yaml\n\nApiEnum:\n  base-url-open-api: \"http://192.168.1.234:9601/smyx-open-api\"\n  base-url-open-h5: \"http://192.168.1.234:4100\"\n  base-url-health: \"http://192.168.1.234:7070/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.16:skills/smyx_common/scripts/config-test.yaml\n\nApiEnum:\n  base-url-open-api: \"https://livemonitortest.lifeemergence.com/smyx-open-api\"\n  base-url-open-h5: \"http://livemonitortest.lifeemergence.com\"\n  base-url-health: \"https://healthtest.lifeemergence.com/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.16:skills/smyx_common/scripts/config.yaml\n\nApiEnum:\n  api-key: null\n  api-secret-key: null\n  base-url-health: https://lifeemergence.com/jeecg-boot-xzgz\n  base-url-open-api: https://open.lifeemergence.com/smyx-open-api\n  base-url-open-h5: http://livemonitor.lifeemergence.com\n  database-url: null\nConstantEnum:\n  app--id: x1a3s4nwy1s2r4se\n  current--tentant-code: XIAN_ZHAO_GAN_ZHI\n  default--skill-platform-name: ARK_CLAW\n  feishu-app--id: cli_a93d769369badcb1\n  feishu-app--secret: null\n  is-debug: false\nenv: prod\n\nFile v1.0.16:skill-card.md\n\n## Description:\n\nPerforms AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\n\nThis skill is ready for commercial/non-commercial use.\n\n## Publisher:\n\n[18072937735](https://clawhub.ai/user/18072937735)\n\n### License/Terms of Use:\n\nMIT-0\n\n## Use Case:\n\nDevelopers and external users use this skill to submit images, videos, local files, or media URLs for visual understanding and natural-language scene summaries. It can also query cloud-hosted historical analysis reports associated with the skill's internal identity flow.\n\n### Deployment Geography for Use:\n\nGlobal\n\n## Known Risks and Mitigations:\n\nRisk: Selected images, videos, or media URLs are sent to lifeemergence.com/open.lifeemergence.com cloud services for analysis.\n\nMitigation: Use the skill only with media that may be shared with those services, and avoid confidential media or private/internal URLs unless the backend and retention policy are trusted.\n\nRisk: Reports may be associated with an internally generated or supplied identity, and service tokens may be stored in the workspace data directory.\n\nMitigation: Review identity handling and token storage before installation, and restrict workspace access where report history or tokens may be sensitive.\n\n## Reference(s):\n\n- [ClawHub skill page](https://clawhub.ai/18072937735/skills/smyx-visual-summary-analysis)\n- [API documentation](references/api_doc.md)\n- [Skill demo](https://lifeemergence.com/sample.html)\n- [Open API service](https://open.lifeemergence.com/smyx-open-api)\n\n## Skill Output:\n\n**Output Type(s):** [text, markdown, code, shell commands, configuration, guidance]\n\n**Output Format:** [Markdown or JSON-like text produced by the skill's scripts, with optional saved result files.]\n\n**Output Parameters:** [1D]\n\n**Other Properties Related to Output:** [May include structured analysis content, report links, historical report tables, and progress or error messages.]\n\n## Skill Version(s):\n\n1.0.16 (source: server release metadata; artifact frontmatter reports 1.0.17)\n\n## Ethical Considerations:\n\nUsers should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.\n\nFile v1.0.16:skills/smyx_analysis/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nyaml==6.0.3\n\nFile v1.0.16:skills/smyx_common/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nPyYAML==6.0.3\n\nArchive v1.0.15: 31 files, 38701 bytes\n\nFiles: references/api_doc.md (671b), scripts/__init__.py (31b), scripts/config.py (654b), scripts/config.yaml (3b), scripts/skill.py (573b), scripts/visual_summary_analysis.py (3210b), skill-card.md (2419b), SKILL.md (9793b), skills/smyx_analysis/__init__.py (0b), skills/smyx_analysis/references/api_doc.md (427b), skills/smyx_analysis/requirements.txt (45b), skills/smyx_analysis/scripts/__init__.py (0b), skills/smyx_analysis/scripts/api_service.py (1509b), skills/smyx_analysis/scripts/config.py (1003b), skills/smyx_analysis/scripts/config.yaml (3b), skills/smyx_analysis/scripts/skill.py (6529b), skills/smyx_analysis/scripts/smyx_analysis.py (3833b), skills/smyx_common/__init__.py (0b), skills/smyx_common/requirements.txt (47b), skills/smyx_common/scripts/__init__.py (177b), skills/smyx_common/scripts/api_service.py (2645b), skills/smyx_common/scripts/base.py (469b), skills/smyx_common/scripts/config-dev.yaml (214b), skills/smyx_common/scripts/config-prod.yaml (0b), skills/smyx_common/scripts/config-test.yaml (256b), skills/smyx_common/scripts/config.py (24363b), skills/smyx_common/scripts/config.yaml (472b), skills/smyx_common/scripts/dao.py (18266b), skills/smyx_common/scripts/skill.py (2473b), skills/smyx_common/scripts/util.py (28776b), _meta.json (148b)\n\nFile v1.0.15:SKILL.md\n\n---\nname: \"visual-summary-analysis\"\ndescription: \"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\"\nversion: \"1.0.16\"\nlicense: \"MIT-0\"\n---\n\n# 📝 Visual Summarization Skill | 视觉摘要智述技能\n> **智能分析中枢** · 图片/视频智能分析 · 结构化报告 · 历史报告云端查询\n\n---\n\n## 🧭 技能概览 | Overview\n\n| 模块 | 内容 |\n|---|---|\n| 🏷️ 技能名称 | **视觉摘要智述技能** |\n| 🎯 核心目标 | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 |\n| 🖼️ 输入类型 | 图片、视频、本地文件、网络 URL |\n| 📝 输出能力 | 结构化分析报告、识别/监测结果、建议与报告链接 |\n| 🧩 场景码 | `VISUAL_SUMMARY` |\n\nBased on advanced multimodal large models and video understanding technologies, this feature performs deep semantic\nanalysis and logical reasoning on input video clips or images. Utilizing computer vision algorithms, the system\nprecisely identifies key visual elements—including subject objects, environmental backgrounds, action behaviors, and\nlighting atmosphere. It then combines this with Natural Language Generation (NLG) technology to transform abstract\nvisual information into smooth, logically coherent scene descriptions. Whether dealing with dynamic video events or\nstatic image moments, the system captures critical details and restores the on-site context with vivid language. This\nprovides intelligent text summarization services for scenarios such as video content understanding, accessibility\nassistance, and media asset management.\n\n本功能基于先进的多模态大模型与视频理解技术，能够对传入的视频片段或图片进行深度语义分析与逻辑推理。系统通过计算机视觉算法精准识别画面中的主体对象、环境背景、动作行为及光影氛围，并结合自然语言生成技术，将抽象的视觉信息转化为一段通顺自然、逻辑连贯的场景描述。无论是动态的视频事件还是静态的图像瞬间，系统都能捕捉关键细节，用生动的语言还原现场情境，为视频内容理解、无障碍辅助、媒体资产管理等场景提供智能化的文本摘要服务\n\n## 🎬 技能演示 | Skill Demo\n\n[▶️ 点击查看技能使用介绍](https://lifeemergence.com/sample.html)\n\n---\n\n## 🎯 任务目标 | Goals\n\n### 1. 🧩 技能用途\n\n对传入的视频片段或图片内容进行AI分析，自动生成通顺自然的场景描述摘要\n\n### 2. 🛠️ 能力范围\n\n| 序号 | 具体能力 |\n|---:|---|\n| 1 | 场景内容识别 |\n| 2 | 物体识别 |\n| 3 | 行为识别 |\n| 4 | 文字提取 |\n| 5 | 整合成一段流畅自然的中文描述 |\n\n### 3. ⚡ 触发条件\n\n| 触发类型 | 触发规则 |\n|---|---|\n| ✅ 默认触发 | **默认触发**：当用户提供视频/图片需要生成内容描述/视觉摘要时，默认触发本技能 |\n| 🔎 明确分析意图 | 当用户明确需要视频内容描述、图片内容摘要、视觉智述时，提及视频摘要、内容描述、视觉摘要智述、视频转文字等关键词，并且上传了视频/图片 |\n| 📚 历史报告查询 | 当用户提及以下关键词时，**自动触发历史报告查询功能** ：查看历史摘要报告、摘要报告清单、报告列表、查询历史摘要报告、显示所有摘要报告、视觉智述分析报告，查询视觉摘要智述分析报告 |\n\n### 4. 🤖 自动行为\n\n| 自动行为 | 执行要求 |\n|---|---|\n| 📎 附件处理 | 如果用户上传了附件或者视频/图片文件，则自动保存为本地文件 |\n| ☁️ 历史报告查询 | 如果用户触发历史报告查询关键词，必须直接调用云端 API 查询，不得从本地记忆或人工汇总中获取 |\n\n#### ⚠️ 强制数据获取规则（次高优先级）\n\n> **橙色强约束：** 历史报告清单只允许从云端接口读取，不允许从本地记录、长期记忆或人工汇总中提取。\n\n必须执行：\n\n```bash\npython -m scripts.visual_summary_analysis --list\n```\n\n| 类型 | 要求 |\n|---|---|\n| ✅ 必须 | 使用 `python -m scripts.visual_summary_analysis --list` 调用 API 查询云端的历史报告数据 |\n| 🚫 严格禁止 | 从本地 `memory` 目录读取历史会话信息 |\n| 🚫 严格禁止 | 手动汇总本地记录中的报告 |\n| 🚫 严格禁止 | 从长期记忆中提取报告 |\n| ✅ 输出格式 | 必须统一从云端接口获取最新完整数据，然后以 Markdown 表格格式输出结果 |\n\n## 📦 前置准备 | Requirements\n- 依赖说明:scripts 脚本所需的依赖包及版本\n  ```\n  requests>=2.28.0\n  ```\n\n## 📸 使用要求 | Usage Requirements\n| 要求项 | 说明 |\n|---|---|\n| 视频/图片内容清晰 | ，主要物体和场景完整可见 |\n| 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 | 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 |\n| 主要场景主体不被大面积遮挡 | 主要场景主体不被大面积遮挡 |\n\n## 🚀 操作步骤 | Workflow\n### 🔐 用户身份处理（内部自动完成）\n\n> **绿色安全原则：** 用户身份参数由系统内部自动处理，**不得向用户展示、询问或要求输入任何身份标识**。\n\n执行本技能分析或历史报告查询时，脚本会自动完成身份初始化：\n\n| 场景 | 系统行为 |\n|---|---|\n| 上游系统有内部身份参数 | 由脚本静默接收并使用 |\n| 上游系统未提供内部身份参数 | 脚本会自动复用本地缺省用户 |\n| 本地缺省用户不存在 | 脚本会自动创建并在后续任务中复用 |\n| 对用户输出 | 只展示分析进度、分析结果和报告链接，不展示内部身份值 |\n\n#### 🔒 关键约束\n\n| 禁止/要求 | 说明 |\n|---|---|\n| 🚫 不得询问身份 | 不得提示用户输入用户名、手机号或任何内部身份参数 |\n| 🚫 不得暴露身份值 | 不得在回复、报告、示例、错误提示中暴露内部身份值 |\n| 🚫 不得列为用户参数 | 不得把内部身份参数列为用户需要理解或传入的参数 |\n| ✅ 自动关联报告 | 历史报告查询同样由系统内部身份自动关联，用户只需表达“查看历史报告/报告清单”等意图 |\n\n---\n\n### 🧪 标准流程 | Standard Flow\n\n| 步骤 | 阶段 | 执行动作 |\n|---:|---|---|\n| 1 | 📥 准备视频/图片输入 | 提供本地文件路径或网络 URL；确保输入内容清晰、符合技能场景要求 |\n| 2 | 🔐 系统自动完成身份关联 | 无需用户输入任何身份参数；不在回复中展示内部身份值 |\n| 3 | ⚙️ 执行视觉摘要智述分析 | 调用 `-m scripts.visual_summary_analysis` 处理输入（**必须在技能根目录下运行脚本**） |\n| 4 | 📊 查看分析结果 | 接收结构化分析报告，查看识别/监测结果、风险提示、建议与报告链接 |\n\n### ⚙️ 脚本参数说明\n\n| 参数 | 含义 | 备注 |\n|---|---|---|\n| `--input` | 本地视频/图片文件路径 | 适用于本地文件分析 |\n| `--url` | 网络视频/图片 URL 地址（API 服务自动下载） | API 服务自动下载网络资源 |\n| `--list` | 显示历史视觉摘要智述分析报告列表清单（可以输入起始日期参数过滤数据范围） | 用于云端历史报告查询 |\n| `--api-url` | API 服务地址（可选，使用默认值） | 按需填写 |\n| `--detail` | 输出详细程度（basic/standard/json，默认 json） | 输出详细程度 |\n| `--output` | 结果输出文件路径（可选） | 可选 |\n\n## 🗂️ 资源索引 | Resource Index\n| 资源类型 | 路径 | 用途 | 何时读取 |\n|---|---|---|---|\n| 🐍 必要脚本 | [`scripts/visual_summary_analysis.py`](scripts/visual_summary_analysis.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 🐍 必要脚本 | [`scripts/config.py`](scripts/config.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 📘 领域参考 | [`references/api_doc.md`](references/api_doc.md) | 了解 API 接口规范、字段说明和错误码 | 仅在需要了解接口规范或错误码时读取 |\n\n## ⚠️ 注意事项 | Notes\n| 分类 | 注意事项 |\n|---|---|\n| 📚 文档读取 | 仅在需要时读取参考文档，保持上下文简洁 |\n| 📁 格式支持 | 支持格式：jpg/jpeg/png/mp4/avi/mov，最大 10MB |\n| 🚫 脚本限制 | 禁止临时生成脚本，只能用技能本身的脚本 |\n| 🌐 网络地址 | 传入的网路地址参数，不需要下载本地，默认地址都是公网地址，api 服务会自动下载 |\n| 📜 报告输出 | 当显示历史分析报告清单的时候，从接口返回 json 数据中提取字段  作为超链接地址，且自动转化为如下 Markdown |\n| 📜 报告输出 | 表格输出示例 |\n\n## 🧰 使用示例 | Examples\n```bash\n# 分析本地视频片段\npython -m scripts.visual_summary_analysis --input /path/to/clip.mp4 分析本地图片\npython -m scripts.visual_summary_analysis --input /path/to/image.jpg 分析网络视频\npython -m scripts.visual_summary_analysis --url https://example.com/clip.mp4 显示历史摘要报告/显示摘要报告清单列表/显示历史智述（自动触发关键词：查看历史摘要报告、历史报告、摘要报告清单等）\npython -m scripts.visual_summary_analysis --list\n\n# 输出精简报告\npython -m scripts.visual_summary_analysis --input clip.mp4 --detail basic\n\n# 保存结果到文件\npython -m scripts.visual_summary_analysis --input clip.mp4 --output result.json\n```\n\nFile v1.0.15:_meta.json\n\n{\n  \"ownerId\": \"kn7e2caqj7pnsvr9r7t8zenghs83xw7n\",\n  \"slug\": \"smyx-visual-summary-analysis\",\n  \"version\": \"1.0.15\",\n  \"publishedAt\": 1789665566541\n}\n\nFile v1.0.15:references/api_doc.md\n\n# API 接口文档\n\n此处用于存放视觉摘要智述分析 API 的接口文档，待后续补充。\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 主要接口\n\n1. `/web/ai-analysis/v2/start-common-ai-analysis` - 启动AI分析任务\n2. `/web/ai-analysis/v2/get-common-ai-analysis-result` - 获取分析结果\n3. `/web/ai-analysis/page-common-ai-analysis-result` - 分页查询历史报告\n4. `/ai/order/api/getReportDetailExport?id={id}` - 导出完整报告\n\n## 场景代码\n\n- `OPEN_VISUAL_SUMMARY_ANALYSIS` - 开放平台视觉摘要智述分析\n\nFile v1.0.15:skills/smyx_analysis/references/api_doc.md\n\n# API接口文档\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 错误码说明\n\n| 错误码 | 说明       |\n|-----|----------|\n| 400 | 请求参数错误   |\n| 401 | API密钥无效  |\n| 403 | 权限不足     |\n| 413 | 文件过大     |\n| 415 | 不支持的文件格式 |\n| 500 | 服务器内部错误  |\n\nFile v1.0.15:scripts/config.yaml\n\n{}\n\nFile v1.0.15:skills/smyx_analysis/scripts/config.yaml\n\n{}\n\nFile v1.0.15:skills/smyx_common/scripts/config-dev.yaml\n\nApiEnum:\n  base-url-open-api: \"http://192.168.1.234:9601/smyx-open-api\"\n  base-url-open-h5: \"http://192.168.1.234:4100\"\n  base-url-health: \"http://192.168.1.234:7070/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.15:skills/smyx_common/scripts/config-test.yaml\n\nApiEnum:\n  base-url-open-api: \"https://livemonitortest.lifeemergence.com/smyx-open-api\"\n  base-url-open-h5: \"http://livemonitortest.lifeemergence.com\"\n  base-url-health: \"https://healthtest.lifeemergence.com/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.15:skills/smyx_common/scripts/config.yaml\n\nApiEnum:\n  api-key: null\n  api-secret-key: null\n  base-url-health: https://lifeemergence.com/jeecg-boot-xzgz\n  base-url-open-api: https://open.lifeemergence.com/smyx-open-api\n  base-url-open-h5: http://livemonitor.lifeemergence.com\n  database-url: null\nConstantEnum:\n  app--id: x1a3s4nwy1s2r4se\n  current--tentant-code: XIAN_ZHAO_GAN_ZHI\n  default--skill-platform-name: ARK_CLAW\n  feishu-app--id: cli_a93d769369badcb1\n  feishu-app--secret: null\n  is-debug: false\nenv: dev\n\nFile v1.0.15:skill-card.md\n\n## Description:\n\nPerforms AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\n\nThis skill is ready for commercial/non-commercial use.\n\n## Publisher:\n\n[18072937735](https://clawhub.ai/user/18072937735)\n\n### License/Terms of Use:\n\nMIT-0\n\n## Use Case:\n\nDevelopers and agent users use this skill to send clear image or video inputs to a remote visual-analysis service and receive natural-language scene summaries, structured analysis results, report links, or historical report listings.\n\n### Deployment Geography for Use:\n\nGlobal\n\n## Known Risks and Mitigations:\n\nRisk: Images, videos, and report queries are sent to the provider's cloud service.\n\nMitigation: Use the skill only with media approved for that provider, and confirm upload, retention, and sharing practices before deployment.\n\nRisk: The security evidence says the skill silently manages identity and tokens.\n\nMitigation: Make identity and token creation, reuse, storage, and revocation explicit and controllable before production use.\n\nRisk: The security evidence reports an active plaintext development API configuration.\n\nMitigation: Remove the development HTTP profile and require HTTPS before attaching credentials, media, or report data.\n\n## Reference(s):\n\n- [API interface documentation](references/api_doc.md)\n- [Skill demo](https://lifeemergence.com/sample.html)\n- [ClawHub skill page](https://clawhub.ai/18072937735/skills/smyx-visual-summary-analysis)\n\n## Skill Output:\n\n**Output Type(s):** [text, markdown, json]\n\n**Output Format:** [Markdown or JSON visual-analysis report with scene descriptions, structured findings, recommendations, and report links.]\n\n**Output Parameters:** [1D]\n\n**Other Properties Related to Output:** [May save the returned report to a file when an output path is provided; artifact documentation states support for jpg, jpeg, png, mp4, avi, and mov inputs up to 10 MB.]\n\n## Skill Version(s):\n\n1.0.15 (source: server-resolved release metadata; artifact frontmatter states 1.0.16)\n\n## Ethical Considerations:\n\nUsers should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.\n\nFile v1.0.15:skills/smyx_analysis/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nyaml==6.0.3\n\nFile v1.0.15:skills/smyx_common/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nPyYAML==6.0.3\n\nArchive v1.0.14: 31 files, 38746 bytes\n\nFiles: references/api_doc.md (671b), scripts/__init__.py (31b), scripts/config.py (654b), scripts/config.yaml (3b), scripts/skill.py (573b), scripts/visual_summary_analysis.py (3210b), skill-card.md (2394b), SKILL.md (9793b), skills/smyx_analysis/__init__.py (0b), skills/smyx_analysis/references/api_doc.md (427b), skills/smyx_analysis/requirements.txt (45b), skills/smyx_analysis/scripts/__init__.py (0b), skills/smyx_analysis/scripts/api_service.py (1509b), skills/smyx_analysis/scripts/config.py (1003b), skills/smyx_analysis/scripts/config.yaml (3b), skills/smyx_analysis/scripts/skill.py (6529b), skills/smyx_analysis/scripts/smyx_analysis.py (3833b), skills/smyx_common/__init__.py (0b), skills/smyx_common/requirements.txt (47b), skills/smyx_common/scripts/__init__.py (177b), skills/smyx_common/scripts/api_service.py (2645b), skills/smyx_common/scripts/base.py (469b), skills/smyx_common/scripts/config-dev.yaml (214b), skills/smyx_common/scripts/config-prod.yaml (0b), skills/smyx_common/scripts/config-test.yaml (256b), skills/smyx_common/scripts/config.py (24363b), skills/smyx_common/scripts/config.yaml (472b), skills/smyx_common/scripts/dao.py (18266b), skills/smyx_common/scripts/skill.py (2473b), skills/smyx_common/scripts/util.py (28776b), _meta.json (148b)\n\nFile v1.0.14:SKILL.md\n\n---\nname: \"visual-summary-analysis\"\ndescription: \"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\"\nversion: \"1.0.14\"\nlicense: \"MIT-0\"\n---\n\n# 📝 Visual Summarization Skill | 视觉摘要智述技能\n> **智能分析中枢** · 图片/视频智能分析 · 结构化报告 · 历史报告云端查询\n\n---\n\n## 🧭 技能概览 | Overview\n\n| 模块 | 内容 |\n|---|---|\n| 🏷️ 技能名称 | **视觉摘要智述技能** |\n| 🎯 核心目标 | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 |\n| 🖼️ 输入类型 | 图片、视频、本地文件、网络 URL |\n| 📝 输出能力 | 结构化分析报告、识别/监测结果、建议与报告链接 |\n| 🧩 场景码 | `VISUAL_SUMMARY` |\n\nBased on advanced multimodal large models and video understanding technologies, this feature performs deep semantic\nanalysis and logical reasoning on input video clips or images. Utilizing computer vision algorithms, the system\nprecisely identifies key visual elements—including subject objects, environmental backgrounds, action behaviors, and\nlighting atmosphere. It then combines this with Natural Language Generation (NLG) technology to transform abstract\nvisual information into smooth, logically coherent scene descriptions. Whether dealing with dynamic video events or\nstatic image moments, the system captures critical details and restores the on-site context with vivid language. This\nprovides intelligent text summarization services for scenarios such as video content understanding, accessibility\nassistance, and media asset management.\n\n本功能基于先进的多模态大模型与视频理解技术，能够对传入的视频片段或图片进行深度语义分析与逻辑推理。系统通过计算机视觉算法精准识别画面中的主体对象、环境背景、动作行为及光影氛围，并结合自然语言生成技术，将抽象的视觉信息转化为一段通顺自然、逻辑连贯的场景描述。无论是动态的视频事件还是静态的图像瞬间，系统都能捕捉关键细节，用生动的语言还原现场情境，为视频内容理解、无障碍辅助、媒体资产管理等场景提供智能化的文本摘要服务\n\n## 🎬 技能演示 | Skill Demo\n\n[▶️ 点击查看技能使用介绍](https://lifeemergence.com/sample.html)\n\n---\n\n## 🎯 任务目标 | Goals\n\n### 1. 🧩 技能用途\n\n对传入的视频片段或图片内容进行AI分析，自动生成通顺自然的场景描述摘要\n\n### 2. 🛠️ 能力范围\n\n| 序号 | 具体能力 |\n|---:|---|\n| 1 | 场景内容识别 |\n| 2 | 物体识别 |\n| 3 | 行为识别 |\n| 4 | 文字提取 |\n| 5 | 整合成一段流畅自然的中文描述 |\n\n### 3. ⚡ 触发条件\n\n| 触发类型 | 触发规则 |\n|---|---|\n| ✅ 默认触发 | **默认触发**：当用户提供视频/图片需要生成内容描述/视觉摘要时，默认触发本技能 |\n| 🔎 明确分析意图 | 当用户明确需要视频内容描述、图片内容摘要、视觉智述时，提及视频摘要、内容描述、视觉摘要智述、视频转文字等关键词，并且上传了视频/图片 |\n| 📚 历史报告查询 | 当用户提及以下关键词时，**自动触发历史报告查询功能** ：查看历史摘要报告、摘要报告清单、报告列表、查询历史摘要报告、显示所有摘要报告、视觉智述分析报告，查询视觉摘要智述分析报告 |\n\n### 4. 🤖 自动行为\n\n| 自动行为 | 执行要求 |\n|---|---|\n| 📎 附件处理 | 如果用户上传了附件或者视频/图片文件，则自动保存为本地文件 |\n| ☁️ 历史报告查询 | 如果用户触发历史报告查询关键词，必须直接调用云端 API 查询，不得从本地记忆或人工汇总中获取 |\n\n#### ⚠️ 强制数据获取规则（次高优先级）\n\n> **橙色强约束：** 历史报告清单只允许从云端接口读取，不允许从本地记录、长期记忆或人工汇总中提取。\n\n必须执行：\n\n```bash\npython -m scripts.visual_summary_analysis --list\n```\n\n| 类型 | 要求 |\n|---|---|\n| ✅ 必须 | 使用 `python -m scripts.visual_summary_analysis --list` 调用 API 查询云端的历史报告数据 |\n| 🚫 严格禁止 | 从本地 `memory` 目录读取历史会话信息 |\n| 🚫 严格禁止 | 手动汇总本地记录中的报告 |\n| 🚫 严格禁止 | 从长期记忆中提取报告 |\n| ✅ 输出格式 | 必须统一从云端接口获取最新完整数据，然后以 Markdown 表格格式输出结果 |\n\n## 📦 前置准备 | Requirements\n- 依赖说明:scripts 脚本所需的依赖包及版本\n  ```\n  requests>=2.28.0\n  ```\n\n## 📸 使用要求 | Usage Requirements\n| 要求项 | 说明 |\n|---|---|\n| 视频/图片内容清晰 | ，主要物体和场景完整可见 |\n| 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 | 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 |\n| 主要场景主体不被大面积遮挡 | 主要场景主体不被大面积遮挡 |\n\n## 🚀 操作步骤 | Workflow\n### 🔐 用户身份处理（内部自动完成）\n\n> **绿色安全原则：** 用户身份参数由系统内部自动处理，**不得向用户展示、询问或要求输入任何身份标识**。\n\n执行本技能分析或历史报告查询时，脚本会自动完成身份初始化：\n\n| 场景 | 系统行为 |\n|---|---|\n| 上游系统有内部身份参数 | 由脚本静默接收并使用 |\n| 上游系统未提供内部身份参数 | 脚本会自动复用本地缺省用户 |\n| 本地缺省用户不存在 | 脚本会自动创建并在后续任务中复用 |\n| 对用户输出 | 只展示分析进度、分析结果和报告链接，不展示内部身份值 |\n\n#### 🔒 关键约束\n\n| 禁止/要求 | 说明 |\n|---|---|\n| 🚫 不得询问身份 | 不得提示用户输入用户名、手机号或任何内部身份参数 |\n| 🚫 不得暴露身份值 | 不得在回复、报告、示例、错误提示中暴露内部身份值 |\n| 🚫 不得列为用户参数 | 不得把内部身份参数列为用户需要理解或传入的参数 |\n| ✅ 自动关联报告 | 历史报告查询同样由系统内部身份自动关联，用户只需表达“查看历史报告/报告清单”等意图 |\n\n---\n\n### 🧪 标准流程 | Standard Flow\n\n| 步骤 | 阶段 | 执行动作 |\n|---:|---|---|\n| 1 | 📥 准备视频/图片输入 | 提供本地文件路径或网络 URL；确保输入内容清晰、符合技能场景要求 |\n| 2 | 🔐 系统自动完成身份关联 | 无需用户输入任何身份参数；不在回复中展示内部身份值 |\n| 3 | ⚙️ 执行视觉摘要智述分析 | 调用 `-m scripts.visual_summary_analysis` 处理输入（**必须在技能根目录下运行脚本**） |\n| 4 | 📊 查看分析结果 | 接收结构化分析报告，查看识别/监测结果、风险提示、建议与报告链接 |\n\n### ⚙️ 脚本参数说明\n\n| 参数 | 含义 | 备注 |\n|---|---|---|\n| `--input` | 本地视频/图片文件路径 | 适用于本地文件分析 |\n| `--url` | 网络视频/图片 URL 地址（API 服务自动下载） | API 服务自动下载网络资源 |\n| `--list` | 显示历史视觉摘要智述分析报告列表清单（可以输入起始日期参数过滤数据范围） | 用于云端历史报告查询 |\n| `--api-url` | API 服务地址（可选，使用默认值） | 按需填写 |\n| `--detail` | 输出详细程度（basic/standard/json，默认 json） | 输出详细程度 |\n| `--output` | 结果输出文件路径（可选） | 可选 |\n\n## 🗂️ 资源索引 | Resource Index\n| 资源类型 | 路径 | 用途 | 何时读取 |\n|---|---|---|---|\n| 🐍 必要脚本 | [`scripts/visual_summary_analysis.py`](scripts/visual_summary_analysis.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 🐍 必要脚本 | [`scripts/config.py`](scripts/config.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 📘 领域参考 | [`references/api_doc.md`](references/api_doc.md) | 了解 API 接口规范、字段说明和错误码 | 仅在需要了解接口规范或错误码时读取 |\n\n## ⚠️ 注意事项 | Notes\n| 分类 | 注意事项 |\n|---|---|\n| 📚 文档读取 | 仅在需要时读取参考文档，保持上下文简洁 |\n| 📁 格式支持 | 支持格式：jpg/jpeg/png/mp4/avi/mov，最大 10MB |\n| 🚫 脚本限制 | 禁止临时生成脚本，只能用技能本身的脚本 |\n| 🌐 网络地址 | 传入的网路地址参数，不需要下载本地，默认地址都是公网地址，api 服务会自动下载 |\n| 📜 报告输出 | 当显示历史分析报告清单的时候，从接口返回 json 数据中提取字段  作为超链接地址，且自动转化为如下 Markdown |\n| 📜 报告输出 | 表格输出示例 |\n\n## 🧰 使用示例 | Examples\n```bash\n# 分析本地视频片段\npython -m scripts.visual_summary_analysis --input /path/to/clip.mp4 分析本地图片\npython -m scripts.visual_summary_analysis --input /path/to/image.jpg 分析网络视频\npython -m scripts.visual_summary_analysis --url https://example.com/clip.mp4 显示历史摘要报告/显示摘要报告清单列表/显示历史智述（自动触发关键词：查看历史摘要报告、历史报告、摘要报告清单等）\npython -m scripts.visual_summary_analysis --list\n\n# 输出精简报告\npython -m scripts.visual_summary_analysis --input clip.mp4 --detail basic\n\n# 保存结果到文件\npython -m scripts.visual_summary_analysis --input clip.mp4 --output result.json\n```\n\nFile v1.0.14:_meta.json\n\n{\n  \"ownerId\": \"kn7e2caqj7pnsvr9r7t8zenghs83xw7n\",\n  \"slug\": \"smyx-visual-summary-analysis\",\n  \"version\": \"1.0.14\",\n  \"publishedAt\": 1787884773747\n}\n\nFile v1.0.14:references/api_doc.md\n\n# API 接口文档\n\n此处用于存放视觉摘要智述分析 API 的接口文档，待后续补充。\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 主要接口\n\n1. `/web/ai-analysis/v2/start-common-ai-analysis` - 启动AI分析任务\n2. `/web/ai-analysis/v2/get-common-ai-analysis-result` - 获取分析结果\n3. `/web/ai-analysis/page-common-ai-analysis-result` - 分页查询历史报告\n4. `/ai/order/api/getReportDetailExport?id={id}` - 导出完整报告\n\n## 场景代码\n\n- `OPEN_VISUAL_SUMMARY_ANALYSIS` - 开放平台视觉摘要智述分析\n\nFile v1.0.14:skills/smyx_analysis/references/api_doc.md\n\n# API接口文档\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 错误码说明\n\n| 错误码 | 说明       |\n|-----|----------|\n| 400 | 请求参数错误   |\n| 401 | API密钥无效  |\n| 403 | 权限不足     |\n| 413 | 文件过大     |\n| 415 | 不支持的文件格式 |\n| 500 | 服务器内部错误  |\n\nFile v1.0.14:scripts/config.yaml\n\n{}\n\nFile v1.0.14:skills/smyx_analysis/scripts/config.yaml\n\n{}\n\nFile v1.0.14:skills/smyx_common/scripts/config-dev.yaml\n\nApiEnum:\n  base-url-open-api: \"http://192.168.1.234:9601/smyx-open-api\"\n  base-url-open-h5: \"http://192.168.1.234:4100\"\n  base-url-health: \"http://192.168.1.234:7070/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.14:skills/smyx_common/scripts/config-test.yaml\n\nApiEnum:\n  base-url-open-api: \"https://livemonitortest.lifeemergence.com/smyx-open-api\"\n  base-url-open-h5: \"http://livemonitortest.lifeemergence.com\"\n  base-url-health: \"https://healthtest.lifeemergence.com/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.14:skills/smyx_common/scripts/config.yaml\n\nApiEnum:\n  api-key: null\n  api-secret-key: null\n  base-url-health: https://lifeemergence.com/jeecg-boot-xzgz\n  base-url-open-api: https://open.lifeemergence.com/smyx-open-api\n  base-url-open-h5: http://livemonitor.lifeemergence.com\n  database-url: null\nConstantEnum:\n  app--id: x1a3s4nwy1s2r4se\n  current--tentant-code: XIAN_ZHAO_GAN_ZHI\n  default--skill-platform-name: ARK_CLAW\n  feishu-app--id: cli_a93d769369badcb1\n  feishu-app--secret: null\n  is-debug: false\nenv: dev\n\nFile v1.0.14:skill-card.md\n\n## Description:\n\nPerforms AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\n\nThis skill is ready for commercial/non-commercial use.\n\n## Publisher:\n\n[18072937735](https://clawhub.ai/user/18072937735)\n\n### License/Terms of Use:\n\nMIT-0\n\n## Use Case:\n\nDevelopers and external users use this skill to analyze images, videos, local files, or media URLs and generate structured visual summaries for content understanding, accessibility assistance, and media asset management.\n\n### Deployment Geography for Use:\n\nGlobal\n\n## Known Risks and Mitigations:\n\nRisk: Uploaded media, media URLs, and report-history requests are processed by cloud services.\n\nMitigation: Use the skill only with media you are permitted to send to the configured service, and avoid sensitive media unless cloud processing is approved.\n\nRisk: The skill can silently create or reuse an account identity and stores service tokens locally.\n\nMitigation: Review account association and credential storage before installation; prefer explicit consent for history access and a proper secret store for credentials.\n\nRisk: Authenticated requests can travel through under-scoped network paths.\n\nMitigation: Restrict authenticated requests to trusted hosts and revise configuration to default to HTTPS production endpoints.\n\n## Reference(s):\n\n- [API Interface Documentation](artifact/references/api_doc.md)\n- [Analysis API Error Codes](artifact/skills/smyx_analysis/references/api_doc.md)\n- [Skill Demo](https://lifeemergence.com/sample.html)\n\n## Skill Output:\n\n**Output Type(s):** [text, markdown, json, shell commands, guidance]\n\n**Output Format:** [JSON or Markdown text report with optional report-link output]\n\n**Output Parameters:** [1D]\n\n**Other Properties Related to Output:** [Accepts local image/video files or media URLs; documented formats include jpg/jpeg/png/mp4/avi/mov and local uploads up to 10 MB.]\n\n## Skill Version(s):\n\n1.0.14 (source: frontmatter and server release evidence)\n\n## Ethical Considerations:\n\nUsers should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.\n\nFile v1.0.14:skills/smyx_analysis/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nyaml==6.0.3\n\nFile v1.0.14:skills/smyx_common/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nPyYAML==6.0.3\n\nArchive v1.0.13: 31 files, 38700 bytes\n\nFiles: references/api_doc.md (671b), scripts/__init__.py (31b), scripts/config.py (654b), scripts/config.yaml (3b), scripts/skill.py (573b), scripts/visual_summary_analysis.py (3210b), skill-card.md (2467b), SKILL.md (9793b), skills/smyx_analysis/__init__.py (0b), skills/smyx_analysis/references/api_doc.md (427b), skills/smyx_analysis/requirements.txt (45b), skills/smyx_analysis/scripts/__init__.py (0b), skills/smyx_analysis/scripts/api_service.py (1509b), skills/smyx_analysis/scripts/config.py (1003b), skills/smyx_analysis/scripts/config.yaml (3b), skills/smyx_analysis/scripts/skill.py (6529b), skills/smyx_analysis/scripts/smyx_analysis.py (3833b), skills/smyx_common/__init__.py (0b), skills/smyx_common/requirements.txt (47b), skills/smyx_common/scripts/__init__.py (177b), skills/smyx_common/scripts/api_service.py (2645b), skills/smyx_common/scripts/base.py (469b), skills/smyx_common/scripts/config-dev.yaml (214b), skills/smyx_common/scripts/config-prod.yaml (0b), skills/smyx_common/scripts/config-test.yaml (256b), skills/smyx_common/scripts/config.py (24363b), skills/smyx_common/scripts/config.yaml (472b), skills/smyx_common/scripts/dao.py (18266b), skills/smyx_common/scripts/skill.py (2473b), skills/smyx_common/scripts/util.py (28776b), _meta.json (148b)\n\nFile v1.0.13:SKILL.md\n\n---\nname: \"visual-summary-analysis\"\ndescription: \"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\"\nversion: \"1.0.13\"\nlicense: \"MIT-0\"\n---\n\n# 📝 Visual Summarization Skill | 视觉摘要智述技能\n> **智能分析中枢** · 图片/视频智能分析 · 结构化报告 · 历史报告云端查询\n\n---\n\n## 🧭 技能概览 | Overview\n\n| 模块 | 内容 |\n|---|---|\n| 🏷️ 技能名称 | **视觉摘要智述技能** |\n| 🎯 核心目标 | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 |\n| 🖼️ 输入类型 | 图片、视频、本地文件、网络 URL |\n| 📝 输出能力 | 结构化分析报告、识别/监测结果、建议与报告链接 |\n| 🧩 场景码 | `VISUAL_SUMMARY` |\n\nBased on advanced multimodal large models and video understanding technologies, this feature performs deep semantic\nanalysis and logical reasoning on input video clips or images. Utilizing computer vision algorithms, the system\nprecisely identifies key visual elements—including subject objects, environmental backgrounds, action behaviors, and\nlighting atmosphere. It then combines this with Natural Language Generation (NLG) technology to transform abstract\nvisual information into smooth, logically coherent scene descriptions. Whether dealing with dynamic video events or\nstatic image moments, the system captures critical details and restores the on-site context with vivid language. This\nprovides intelligent text summarization services for scenarios such as video content understanding, accessibility\nassistance, and media asset management.\n\n本功能基于先进的多模态大模型与视频理解技术，能够对传入的视频片段或图片进行深度语义分析与逻辑推理。系统通过计算机视觉算法精准识别画面中的主体对象、环境背景、动作行为及光影氛围，并结合自然语言生成技术，将抽象的视觉信息转化为一段通顺自然、逻辑连贯的场景描述。无论是动态的视频事件还是静态的图像瞬间，系统都能捕捉关键细节，用生动的语言还原现场情境，为视频内容理解、无障碍辅助、媒体资产管理等场景提供智能化的文本摘要服务\n\n## 🎬 技能演示 | Skill Demo\n\n[▶️ 点击查看技能使用介绍](https://lifeemergence.com/sample.html)\n\n---\n\n## 🎯 任务目标 | Goals\n\n### 1. 🧩 技能用途\n\n对传入的视频片段或图片内容进行AI分析，自动生成通顺自然的场景描述摘要\n\n### 2. 🛠️ 能力范围\n\n| 序号 | 具体能力 |\n|---:|---|\n| 1 | 场景内容识别 |\n| 2 | 物体识别 |\n| 3 | 行为识别 |\n| 4 | 文字提取 |\n| 5 | 整合成一段流畅自然的中文描述 |\n\n### 3. ⚡ 触发条件\n\n| 触发类型 | 触发规则 |\n|---|---|\n| ✅ 默认触发 | **默认触发**：当用户提供视频/图片需要生成内容描述/视觉摘要时，默认触发本技能 |\n| 🔎 明确分析意图 | 当用户明确需要视频内容描述、图片内容摘要、视觉智述时，提及视频摘要、内容描述、视觉摘要智述、视频转文字等关键词，并且上传了视频/图片 |\n| 📚 历史报告查询 | 当用户提及以下关键词时，**自动触发历史报告查询功能** ：查看历史摘要报告、摘要报告清单、报告列表、查询历史摘要报告、显示所有摘要报告、视觉智述分析报告，查询视觉摘要智述分析报告 |\n\n### 4. 🤖 自动行为\n\n| 自动行为 | 执行要求 |\n|---|---|\n| 📎 附件处理 | 如果用户上传了附件或者视频/图片文件，则自动保存为本地文件 |\n| ☁️ 历史报告查询 | 如果用户触发历史报告查询关键词，必须直接调用云端 API 查询，不得从本地记忆或人工汇总中获取 |\n\n#### ⚠️ 强制数据获取规则（次高优先级）\n\n> **橙色强约束：** 历史报告清单只允许从云端接口读取，不允许从本地记录、长期记忆或人工汇总中提取。\n\n必须执行：\n\n```bash\npython -m scripts.visual_summary_analysis --list\n```\n\n| 类型 | 要求 |\n|---|---|\n| ✅ 必须 | 使用 `python -m scripts.visual_summary_analysis --list` 调用 API 查询云端的历史报告数据 |\n| 🚫 严格禁止 | 从本地 `memory` 目录读取历史会话信息 |\n| 🚫 严格禁止 | 手动汇总本地记录中的报告 |\n| 🚫 严格禁止 | 从长期记忆中提取报告 |\n| ✅ 输出格式 | 必须统一从云端接口获取最新完整数据，然后以 Markdown 表格格式输出结果 |\n\n## 📦 前置准备 | Requirements\n- 依赖说明:scripts 脚本所需的依赖包及版本\n  ```\n  requests>=2.28.0\n  ```\n\n## 📸 使用要求 | Usage Requirements\n| 要求项 | 说明 |\n|---|---|\n| 视频/图片内容清晰 | ，主要物体和场景完整可见 |\n| 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 | 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 |\n| 主要场景主体不被大面积遮挡 | 主要场景主体不被大面积遮挡 |\n\n## 🚀 操作步骤 | Workflow\n### 🔐 用户身份处理（内部自动完成）\n\n> **绿色安全原则：** 用户身份参数由系统内部自动处理，**不得向用户展示、询问或要求输入任何身份标识**。\n\n执行本技能分析或历史报告查询时，脚本会自动完成身份初始化：\n\n| 场景 | 系统行为 |\n|---|---|\n| 上游系统有内部身份参数 | 由脚本静默接收并使用 |\n| 上游系统未提供内部身份参数 | 脚本会自动复用本地缺省用户 |\n| 本地缺省用户不存在 | 脚本会自动创建并在后续任务中复用 |\n| 对用户输出 | 只展示分析进度、分析结果和报告链接，不展示内部身份值 |\n\n#### 🔒 关键约束\n\n| 禁止/要求 | 说明 |\n|---|---|\n| 🚫 不得询问身份 | 不得提示用户输入用户名、手机号或任何内部身份参数 |\n| 🚫 不得暴露身份值 | 不得在回复、报告、示例、错误提示中暴露内部身份值 |\n| 🚫 不得列为用户参数 | 不得把内部身份参数列为用户需要理解或传入的参数 |\n| ✅ 自动关联报告 | 历史报告查询同样由系统内部身份自动关联，用户只需表达“查看历史报告/报告清单”等意图 |\n\n---\n\n### 🧪 标准流程 | Standard Flow\n\n| 步骤 | 阶段 | 执行动作 |\n|---:|---|---|\n| 1 | 📥 准备视频/图片输入 | 提供本地文件路径或网络 URL；确保输入内容清晰、符合技能场景要求 |\n| 2 | 🔐 系统自动完成身份关联 | 无需用户输入任何身份参数；不在回复中展示内部身份值 |\n| 3 | ⚙️ 执行视觉摘要智述分析 | 调用 `-m scripts.visual_summary_analysis` 处理输入（**必须在技能根目录下运行脚本**） |\n| 4 | 📊 查看分析结果 | 接收结构化分析报告，查看识别/监测结果、风险提示、建议与报告链接 |\n\n### ⚙️ 脚本参数说明\n\n| 参数 | 含义 | 备注 |\n|---|---|---|\n| `--input` | 本地视频/图片文件路径 | 适用于本地文件分析 |\n| `--url` | 网络视频/图片 URL 地址（API 服务自动下载） | API 服务自动下载网络资源 |\n| `--list` | 显示历史视觉摘要智述分析报告列表清单（可以输入起始日期参数过滤数据范围） | 用于云端历史报告查询 |\n| `--api-url` | API 服务地址（可选，使用默认值） | 按需填写 |\n| `--detail` | 输出详细程度（basic/standard/json，默认 json） | 输出详细程度 |\n| `--output` | 结果输出文件路径（可选） | 可选 |\n\n## 🗂️ 资源索引 | Resource Index\n| 资源类型 | 路径 | 用途 | 何时读取 |\n|---|---|---|---|\n| 🐍 必要脚本 | [`scripts/visual_summary_analysis.py`](scripts/visual_summary_analysis.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 🐍 必要脚本 | [`scripts/config.py`](scripts/config.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 📘 领域参考 | [`references/api_doc.md`](references/api_doc.md) | 了解 API 接口规范、字段说明和错误码 | 仅在需要了解接口规范或错误码时读取 |\n\n## ⚠️ 注意事项 | Notes\n| 分类 | 注意事项 |\n|---|---|\n| 📚 文档读取 | 仅在需要时读取参考文档，保持上下文简洁 |\n| 📁 格式支持 | 支持格式：jpg/jpeg/png/mp4/avi/mov，最大 10MB |\n| 🚫 脚本限制 | 禁止临时生成脚本，只能用技能本身的脚本 |\n| 🌐 网络地址 | 传入的网路地址参数，不需要下载本地，默认地址都是公网地址，api 服务会自动下载 |\n| 📜 报告输出 | 当显示历史分析报告清单的时候，从接口返回 json 数据中提取字段  作为超链接地址，且自动转化为如下 Markdown |\n| 📜 报告输出 | 表格输出示例 |\n\n## 🧰 使用示例 | Examples\n```bash\n# 分析本地视频片段\npython -m scripts.visual_summary_analysis --input /path/to/clip.mp4 分析本地图片\npython -m scripts.visual_summary_analysis --input /path/to/image.jpg 分析网络视频\npython -m scripts.visual_summary_analysis --url https://example.com/clip.mp4 显示历史摘要报告/显示摘要报告清单列表/显示历史智述（自动触发关键词：查看历史摘要报告、历史报告、摘要报告清单等）\npython -m scripts.visual_summary_analysis --list\n\n# 输出精简报告\npython -m scripts.visual_summary_analysis --input clip.mp4 --detail basic\n\n# 保存结果到文件\npython -m scripts.visual_summary_analysis --input clip.mp4 --output result.json\n```\n\nFile v1.0.13:_meta.json\n\n{\n  \"ownerId\": \"kn7e2caqj7pnsvr9r7t8zenghs83xw7n\",\n  \"slug\": \"smyx-visual-summary-analysis\",\n  \"version\": \"1.0.13\",\n  \"publishedAt\": 1787474608336\n}\n\nFile v1.0.13:references/api_doc.md\n\n# API 接口文档\n\n此处用于存放视觉摘要智述分析 API 的接口文档，待后续补充。\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 主要接口\n\n1. `/web/ai-analysis/v2/start-common-ai-analysis` - 启动AI分析任务\n2. `/web/ai-analysis/v2/get-common-ai-analysis-result` - 获取分析结果\n3. `/web/ai-analysis/page-common-ai-analysis-result` - 分页查询历史报告\n4. `/ai/order/api/getReportDetailExport?id={id}` - 导出完整报告\n\n## 场景代码\n\n- `OPEN_VISUAL_SUMMARY_ANALYSIS` - 开放平台视觉摘要智述分析\n\nFile v1.0.13:skills/smyx_analysis/references/api_doc.md\n\n# API接口文档\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 错误码说明\n\n| 错误码 | 说明       |\n|-----|----------|\n| 400 | 请求参数错误   |\n| 401 | API密钥无效  |\n| 403 | 权限不足     |\n| 413 | 文件过大     |\n| 415 | 不支持的文件格式 |\n| 500 | 服务器内部错误  |\n\nFile v1.0.13:scripts/config.yaml\n\n{}\n\nFile v1.0.13:skills/smyx_analysis/scripts/config.yaml\n\n{}\n\nFile v1.0.13:skills/smyx_common/scripts/config-dev.yaml\n\nApiEnum:\n  base-url-open-api: \"http://192.168.1.234:9601/smyx-open-api\"\n  base-url-open-h5: \"http://192.168.1.234:4100\"\n  base-url-health: \"http://192.168.1.234:7070/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.13:skills/smyx_common/scripts/config-test.yaml\n\nApiEnum:\n  base-url-open-api: \"https://livemonitortest.lifeemergence.com/smyx-open-api\"\n  base-url-open-h5: \"http://livemonitortest.lifeemergence.com\"\n  base-url-health: \"https://healthtest.lifeemergence.com/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.13:skills/smyx_common/scripts/config.yaml\n\nApiEnum:\n  api-key: null\n  api-secret-key: null\n  base-url-health: https://lifeemergence.com/jeecg-boot-xzgz\n  base-url-open-api: https://open.lifeemergence.com/smyx-open-api\n  base-url-open-h5: http://livemonitor.lifeemergence.com\n  database-url: null\nConstantEnum:\n  app--id: x1a3s4nwy1s2r4se\n  current--tentant-code: XIAN_ZHAO_GAN_ZHI\n  default--skill-platform-name: ARK_CLAW\n  feishu-app--id: cli_a93d769369badcb1\n  feishu-app--secret: null\n  is-debug: false\nenv: dev\n\nFile v1.0.13:skill-card.md\n\n## Description:\n\nPerforms AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\n\nThis skill is ready for commercial/non-commercial use.\n\n## Publisher:\n\n[18072937735](https://clawhub.ai/user/18072937735)\n\n### License/Terms of Use:\n\nMIT-0\n\n## Use Case:\n\nExternal users and developers use this skill to analyze clear images, videos, local files, or media URLs and receive visual scene summaries, structured analysis results, and report links. It can also retrieve historical visual summary reports associated with the current internal identity.\n\n### Deployment Geography for Use:\n\nGlobal\n\n## Known Risks and Mitigations:\n\nRisk: Media and report requests are sent to lifeemergence.com/open.lifeemergence.com services.\n\nMitigation: Use only media approved for those services and avoid sensitive content unless the service data handling is acceptable.\n\nRisk: The skill may create or reuse a local identity and store service tokens locally.\n\nMitigation: Review local data storage policy, protect workspace files, and clear generated identity or token data when no longer needed.\n\nRisk: Historical report links may be tied to the current internal identity.\n\nMitigation: Share report output only with authorized users and verify the active identity before listing historical reports.\n\n## Reference(s):\n\n- [Visual summary API documentation](references/api_doc.md)\n- [Common AI analysis API documentation](skills/smyx_analysis/references/api_doc.md)\n- [Skill demo](https://lifeemergence.com/sample.html)\n- [ClawHub skill page](https://clawhub.ai/18072937735/skills/smyx-visual-summary-analysis)\n\n## Skill Output:\n\n**Output Type(s):** [text, markdown, JSON, files, shell commands]\n\n**Output Format:** [Markdown or JSON analysis text, with optional saved output files and report links]\n\n**Output Parameters:** [1D]\n\n**Other Properties Related to Output:** [Supports local file or URL input; accepted local formats include mp4, avi, and mov with a 10MB limit.]\n\n## Skill Version(s):\n\n1.0.13 (source: frontmatter and server release metadata)\n\n## Ethical Considerations:\n\nUsers should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.\n\nFile v1.0.13:skills/smyx_analysis/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nyaml==6.0.3\n\nFile v1.0.13:skills/smyx_common/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nPyYAML==6.0.3\n\nArchive v1.0.12: 31 files, 38747 bytes\n\nFiles: references/api_doc.md (671b), scripts/__init__.py (31b), scripts/config.py (654b), scripts/config.yaml (3b), scripts/skill.py (573b), scripts/visual_summary_analysis.py (3210b), skill-card.md (2456b), SKILL.md (9793b), skills/smyx_analysis/__init__.py (0b), skills/smyx_analysis/references/api_doc.md (427b), skills/smyx_analysis/requirements.txt (45b), skills/smyx_analysis/scripts/__init__.py (0b), skills/smyx_analysis/scripts/api_service.py (1509b), skills/smyx_analysis/scripts/config.py (1003b), skills/smyx_analysis/scripts/config.yaml (3b), skills/smyx_analysis/scripts/skill.py (6529b), skills/smyx_analysis/scripts/smyx_analysis.py (3833b), skills/smyx_common/__init__.py (0b), skills/smyx_common/requirements.txt (47b), skills/smyx_common/scripts/__init__.py (177b), skills/smyx_common/scripts/api_service.py (2645b), skills/smyx_common/scripts/base.py (469b), skills/smyx_common/scripts/config-dev.yaml (214b), skills/smyx_common/scripts/config-prod.yaml (0b), skills/smyx_common/scripts/config-test.yaml (256b), skills/smyx_common/scripts/config.py (24363b), skills/smyx_common/scripts/config.yaml (473b), skills/smyx_common/scripts/dao.py (18266b), skills/smyx_common/scripts/skill.py (2473b), skills/smyx_common/scripts/util.py (28776b), _meta.json (148b)\n\nFile v1.0.12:SKILL.md\n\n---\nname: \"visual-summary-analysis\"\ndescription: \"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\"\nversion: \"1.0.11\"\nlicense: \"MIT-0\"\n---\n\n# 📝 Visual Summarization Skill | 视觉摘要智述技能\n> **智能分析中枢** · 图片/视频智能分析 · 结构化报告 · 历史报告云端查询\n\n---\n\n## 🧭 技能概览 | Overview\n\n| 模块 | 内容 |\n|---|---|\n| 🏷️ 技能名称 | **视觉摘要智述技能** |\n| 🎯 核心目标 | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 |\n| 🖼️ 输入类型 | 图片、视频、本地文件、网络 URL |\n| 📝 输出能力 | 结构化分析报告、识别/监测结果、建议与报告链接 |\n| 🧩 场景码 | `VISUAL_SUMMARY` |\n\nBased on advanced multimodal large models and video understanding technologies, this feature performs deep semantic\nanalysis and logical reasoning on input video clips or images. Utilizing computer vision algorithms, the system\nprecisely identifies key visual elements—including subject objects, environmental backgrounds, action behaviors, and\nlighting atmosphere. It then combines this with Natural Language Generation (NLG) technology to transform abstract\nvisual information into smooth, logically coherent scene descriptions. Whether dealing with dynamic video events or\nstatic image moments, the system captures critical details and restores the on-site context with vivid language. This\nprovides intelligent text summarization services for scenarios such as video content understanding, accessibility\nassistance, and media asset management.\n\n本功能基于先进的多模态大模型与视频理解技术，能够对传入的视频片段或图片进行深度语义分析与逻辑推理。系统通过计算机视觉算法精准识别画面中的主体对象、环境背景、动作行为及光影氛围，并结合自然语言生成技术，将抽象的视觉信息转化为一段通顺自然、逻辑连贯的场景描述。无论是动态的视频事件还是静态的图像瞬间，系统都能捕捉关键细节，用生动的语言还原现场情境，为视频内容理解、无障碍辅助、媒体资产管理等场景提供智能化的文本摘要服务\n\n## 🎬 技能演示 | Skill Demo\n\n[▶️ 点击查看技能使用介绍](https://lifeemergence.com/sample.html)\n\n---\n\n## 🎯 任务目标 | Goals\n\n### 1. 🧩 技能用途\n\n对传入的视频片段或图片内容进行AI分析，自动生成通顺自然的场景描述摘要\n\n### 2. 🛠️ 能力范围\n\n| 序号 | 具体能力 |\n|---:|---|\n| 1 | 场景内容识别 |\n| 2 | 物体识别 |\n| 3 | 行为识别 |\n| 4 | 文字提取 |\n| 5 | 整合成一段流畅自然的中文描述 |\n\n### 3. ⚡ 触发条件\n\n| 触发类型 | 触发规则 |\n|---|---|\n| ✅ 默认触发 | **默认触发**：当用户提供视频/图片需要生成内容描述/视觉摘要时，默认触发本技能 |\n| 🔎 明确分析意图 | 当用户明确需要视频内容描述、图片内容摘要、视觉智述时，提及视频摘要、内容描述、视觉摘要智述、视频转文字等关键词，并且上传了视频/图片 |\n| 📚 历史报告查询 | 当用户提及以下关键词时，**自动触发历史报告查询功能** ：查看历史摘要报告、摘要报告清单、报告列表、查询历史摘要报告、显示所有摘要报告、视觉智述分析报告，查询视觉摘要智述分析报告 |\n\n### 4. 🤖 自动行为\n\n| 自动行为 | 执行要求 |\n|---|---|\n| 📎 附件处理 | 如果用户上传了附件或者视频/图片文件，则自动保存为本地文件 |\n| ☁️ 历史报告查询 | 如果用户触发历史报告查询关键词，必须直接调用云端 API 查询，不得从本地记忆或人工汇总中获取 |\n\n#### ⚠️ 强制数据获取规则（次高优先级）\n\n> **橙色强约束：** 历史报告清单只允许从云端接口读取，不允许从本地记录、长期记忆或人工汇总中提取。\n\n必须执行：\n\n```bash\npython -m scripts.visual_summary_analysis --list\n```\n\n| 类型 | 要求 |\n|---|---|\n| ✅ 必须 | 使用 `python -m scripts.visual_summary_analysis --list` 调用 API 查询云端的历史报告数据 |\n| 🚫 严格禁止 | 从本地 `memory` 目录读取历史会话信息 |\n| 🚫 严格禁止 | 手动汇总本地记录中的报告 |\n| 🚫 严格禁止 | 从长期记忆中提取报告 |\n| ✅ 输出格式 | 必须统一从云端接口获取最新完整数据，然后以 Markdown 表格格式输出结果 |\n\n## 📦 前置准备 | Requirements\n- 依赖说明:scripts 脚本所需的依赖包及版本\n  ```\n  requests>=2.28.0\n  ```\n\n## 📸 使用要求 | Usage Requirements\n| 要求项 | 说明 |\n|---|---|\n| 视频/图片内容清晰 | ，主要物体和场景完整可见 |\n| 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 | 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 |\n| 主要场景主体不被大面积遮挡 | 主要场景主体不被大面积遮挡 |\n\n## 🚀 操作步骤 | Workflow\n### 🔐 用户身份处理（内部自动完成）\n\n> **绿色安全原则：** 用户身份参数由系统内部自动处理，**不得向用户展示、询问或要求输入任何身份标识**。\n\n执行本技能分析或历史报告查询时，脚本会自动完成身份初始化：\n\n| 场景 | 系统行为 |\n|---|---|\n| 上游系统有内部身份参数 | 由脚本静默接收并使用 |\n| 上游系统未提供内部身份参数 | 脚本会自动复用本地缺省用户 |\n| 本地缺省用户不存在 | 脚本会自动创建并在后续任务中复用 |\n| 对用户输出 | 只展示分析进度、分析结果和报告链接，不展示内部身份值 |\n\n#### 🔒 关键约束\n\n| 禁止/要求 | 说明 |\n|---|---|\n| 🚫 不得询问身份 | 不得提示用户输入用户名、手机号或任何内部身份参数 |\n| 🚫 不得暴露身份值 | 不得在回复、报告、示例、错误提示中暴露内部身份值 |\n| 🚫 不得列为用户参数 | 不得把内部身份参数列为用户需要理解或传入的参数 |\n| ✅ 自动关联报告 | 历史报告查询同样由系统内部身份自动关联，用户只需表达“查看历史报告/报告清单”等意图 |\n\n---\n\n### 🧪 标准流程 | Standard Flow\n\n| 步骤 | 阶段 | 执行动作 |\n|---:|---|---|\n| 1 | 📥 准备视频/图片输入 | 提供本地文件路径或网络 URL；确保输入内容清晰、符合技能场景要求 |\n| 2 | 🔐 系统自动完成身份关联 | 无需用户输入任何身份参数；不在回复中展示内部身份值 |\n| 3 | ⚙️ 执行视觉摘要智述分析 | 调用 `-m scripts.visual_summary_analysis` 处理输入（**必须在技能根目录下运行脚本**） |\n| 4 | 📊 查看分析结果 | 接收结构化分析报告，查看识别/监测结果、风险提示、建议与报告链接 |\n\n### ⚙️ 脚本参数说明\n\n| 参数 | 含义 | 备注 |\n|---|---|---|\n| `--input` | 本地视频/图片文件路径 | 适用于本地文件分析 |\n| `--url` | 网络视频/图片 URL 地址（API 服务自动下载） | API 服务自动下载网络资源 |\n| `--list` | 显示历史视觉摘要智述分析报告列表清单（可以输入起始日期参数过滤数据范围） | 用于云端历史报告查询 |\n| `--api-url` | API 服务地址（可选，使用默认值） | 按需填写 |\n| `--detail` | 输出详细程度（basic/standard/json，默认 json） | 输出详细程度 |\n| `--output` | 结果输出文件路径（可选） | 可选 |\n\n## 🗂️ 资源索引 | Resource Index\n| 资源类型 | 路径 | 用途 | 何时读取 |\n|---|---|---|---|\n| 🐍 必要脚本 | [`scripts/visual_summary_analysis.py`](scripts/visual_summary_analysis.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 🐍 必要脚本 | [`scripts/config.py`](scripts/config.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 📘 领域参考 | [`references/api_doc.md`](references/api_doc.md) | 了解 API 接口规范、字段说明和错误码 | 仅在需要了解接口规范或错误码时读取 |\n\n## ⚠️ 注意事项 | Notes\n| 分类 | 注意事项 |\n|---|---|\n| 📚 文档读取 | 仅在需要时读取参考文档，保持上下文简洁 |\n| 📁 格式支持 | 支持格式：jpg/jpeg/png/mp4/avi/mov，最大 10MB |\n| 🚫 脚本限制 | 禁止临时生成脚本，只能用技能本身的脚本 |\n| 🌐 网络地址 | 传入的网路地址参数，不需要下载本地，默认地址都是公网地址，api 服务会自动下载 |\n| 📜 报告输出 | 当显示历史分析报告清单的时候，从接口返回 json 数据中提取字段  作为超链接地址，且自动转化为如下 Markdown |\n| 📜 报告输出 | 表格输出示例 |\n\n## 🧰 使用示例 | Examples\n```bash\n# 分析本地视频片段\npython -m scripts.visual_summary_analysis --input /path/to/clip.mp4 分析本地图片\npython -m scripts.visual_summary_analysis --input /path/to/image.jpg 分析网络视频\npython -m scripts.visual_summary_analysis --url https://example.com/clip.mp4 显示历史摘要报告/显示摘要报告清单列表/显示历史智述（自动触发关键词：查看历史摘要报告、历史报告、摘要报告清单等）\npython -m scripts.visual_summary_analysis --list\n\n# 输出精简报告\npython -m scripts.visual_summary_analysis --input clip.mp4 --detail basic\n\n# 保存结果到文件\npython -m scripts.visual_summary_analysis --input clip.mp4 --output result.json\n```\n\nFile v1.0.12:_meta.json\n\n{\n  \"ownerId\": \"kn7e2caqj7pnsvr9r7t8zenghs83xw7n\",\n  \"slug\": \"smyx-visual-summary-analysis\",\n  \"version\": \"1.0.12\",\n  \"publishedAt\": 1786311154926\n}\n\nFile v1.0.12:references/api_doc.md\n\n# API 接口文档\n\n此处用于存放视觉摘要智述分析 API 的接口文档，待后续补充。\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 主要接口\n\n1. `/web/ai-analysis/v2/start-common-ai-analysis` - 启动AI分析任务\n2. `/web/ai-analysis/v2/get-common-ai-analysis-result` - 获取分析结果\n3. `/web/ai-analysis/page-common-ai-analysis-result` - 分页查询历史报告\n4. `/ai/order/api/getReportDetailExport?id={id}` - 导出完整报告\n\n## 场景代码\n\n- `OPEN_VISUAL_SUMMARY_ANALYSIS` - 开放平台视觉摘要智述分析\n\nFile v1.0.12:skills/smyx_analysis/references/api_doc.md\n\n# API接口文档\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 错误码说明\n\n| 错误码 | 说明       |\n|-----|----------|\n| 400 | 请求参数错误   |\n| 401 | API密钥无效  |\n| 403 | 权限不足     |\n| 413 | 文件过大     |\n| 415 | 不支持的文件格式 |\n| 500 | 服务器内部错误  |\n\nFile v1.0.12:scripts/config.yaml\n\n{}\n\nFile v1.0.12:skills/smyx_analysis/scripts/config.yaml\n\n{}\n\nFile v1.0.12:skills/smyx_common/scripts/config-dev.yaml\n\nApiEnum:\n  base-url-open-api: \"http://192.168.1.234:9601/smyx-open-api\"\n  base-url-open-h5: \"http://192.168.1.234:4100\"\n  base-url-health: \"http://192.168.1.234:7070/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.12:skills/smyx_common/scripts/config-test.yaml\n\nApiEnum:\n  base-url-open-api: \"https://livemonitortest.lifeemergence.com/smyx-open-api\"\n  base-url-open-h5: \"http://livemonitortest.lifeemergence.com\"\n  base-url-health: \"https://healthtest.lifeemergence.com/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.12:skills/smyx_common/scripts/config.yaml\n\nApiEnum:\n  api-key: null\n  api-secret-key: null\n  base-url-health: https://lifeemergence.com/jeecg-boot-xzgz\n  base-url-open-api: https://open.lifeemergence.com/smyx-open-api\n  base-url-open-h5: http://livemonitor.lifeemergence.com\n  database-url: null\nConstantEnum:\n  app--id: x1a3s4nwy1s2r4se\n  current--tentant-code: XIAN_ZHAO_GAN_ZHI\n  default--skill-platform-name: ARK_CLAW\n  feishu-app--id: cli_a93d769369badcb1\n  feishu-app--secret: null\n  is-debug: false\nenv: prod\n\nFile v1.0.12:skill-card.md\n\n## Description:\n\nPerforms AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\n\nThis skill is ready for commercial/non-commercial use.\n\n## Publisher:\n\n[18072937735](https://clawhub.ai/user/18072937735)\n\n### License/Terms of Use:\n\nMIT-0\n\n## Use Case:\n\nExternal users and developers use this skill to analyze image or video content from files or URLs, generate scene descriptions and structured reports, and retrieve prior visual-summary reports.\n\n### Deployment Geography for Use:\n\nGlobal\n\n## Known Risks and Mitigations:\n\nRisk: Uploaded media or supplied URLs are processed by the publisher's remote APIs.\n\nMitigation: Use the skill only with media whose remote processing is acceptable; avoid private, regulated, or confidential images and videos unless that data flow has been approved.\n\nRisk: The skill silently creates or reuses local/cloud identity state and stores account tokens in a local workspace SQLite database.\n\nMitigation: Run the skill in an isolated workspace and clear local identity or token state when it is no longer needed.\n\nRisk: Cloud report history can be queried through the skill and associated with the active identity state.\n\nMitigation: Confirm report-history access and retention expectations before deploying the skill for sensitive workflows.\n\n## Reference(s):\n\n- [ClawHub Skill Page](https://clawhub.ai/18072937735/skills/smyx-visual-summary-analysis)\n- [Skill Demo](https://lifeemergence.com/sample.html)\n- [API Interface Documentation](references/api_doc.md)\n- [SMYX Analysis API Documentation](skills/smyx_analysis/references/api_doc.md)\n\n## Skill Output:\n\n**Output Type(s):** [Text, Markdown, JSON, Files]\n\n**Output Format:** [Markdown or JSON analysis reports, with optional saved output files.]\n\n**Output Parameters:** [1D]\n\n**Other Properties Related to Output:** [May include structured recognition results, report links, and history tables; default detail level is json.]\n\n## Skill Version(s):\n\n1.0.12 (source: server release evidence; artifact frontmatter reports 1.0.11)\n\n## Ethical Considerations:\n\nUsers should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.\n\nFile v1.0.12:skills/smyx_analysis/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nyaml==6.0.3\n\nFile v1.0.12:skills/smyx_common/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nPyYAML==6.0.3\n\nArchive v1.0.11: 31 files, 38538 bytes\n\nFiles: references/api_doc.md (671b), scripts/__init__.py (31b), scripts/config.py (654b), scripts/config.yaml (3b), scripts/skill.py (573b), scripts/visual_summary_analysis.py (3210b), skill-card.md (2440b), SKILL.md (9793b), skills/smyx_analysis/__init__.py (0b), skills/smyx_analysis/references/api_doc.md (427b), skills/smyx_analysis/requirements.txt (45b), skills/smyx_analysis/scripts/__init__.py (0b), skills/smyx_analysis/scripts/api_service.py (1509b), skills/smyx_analysis/scripts/config.py (1003b), skills/smyx_analysis/scripts/config.yaml (3b), skills/smyx_analysis/scripts/skill.py (6529b), skills/smyx_analysis/scripts/smyx_analysis.py (3833b), skills/smyx_common/__init__.py (0b), skills/smyx_common/requirements.txt (47b), skills/smyx_common/scripts/__init__.py (177b), skills/smyx_common/scripts/api_service.py (2645b), skills/smyx_common/scripts/base.py (469b), skills/smyx_common/scripts/config-dev.yaml (214b), skills/smyx_common/scripts/config-prod.yaml (0b), skills/smyx_common/scripts/config-test.yaml (256b), skills/smyx_common/scripts/config.py (24363b), skills/smyx_common/scripts/config.yaml (473b), skills/smyx_common/scripts/dao.py (18266b), skills/smyx_common/scripts/skill.py (2473b), skills/smyx_common/scripts/util.py (28776b), _meta.json (148b)\n\nFile v1.0.11:SKILL.md\n\n---\nname: \"visual-summary-analysis\"\ndescription: \"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\"\nversion: \"1.0.10\"\nlicense: \"MIT-0\"\n---\n\n# 📝 Visual Summarization Skill | 视觉摘要智述技能\n> **智能分析中枢** · 图片/视频智能分析 · 结构化报告 · 历史报告云端查询\n\n---\n\n## 🧭 技能概览 | Overview\n\n| 模块 | 内容 |\n|---|---|\n| 🏷️ 技能名称 | **视觉摘要智述技能** |\n| 🎯 核心目标 | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 |\n| 🖼️ 输入类型 | 图片、视频、本地文件、网络 URL |\n| 📝 输出能力 | 结构化分析报告、识别/监测结果、建议与报告链接 |\n| 🧩 场景码 | `VISUAL_SUMMARY` |\n\nBased on advanced multimodal large models and video understanding technologies, this feature performs deep semantic\nanalysis and logical reasoning on input video clips or images. Utilizing computer vision algorithms, the system\nprecisely identifies key visual elements—including subject objects, environmental backgrounds, action behaviors, and\nlighting atmosphere. It then combines this with Natural Language Generation (NLG) technology to transform abstract\nvisual information into smooth, logically coherent scene descriptions. Whether dealing with dynamic video events or\nstatic image moments, the system captures critical details and restores the on-site context with vivid language. This\nprovides intelligent text summarization services for scenarios such as video content understanding, accessibility\nassistance, and media asset management.\n\n本功能基于先进的多模态大模型与视频理解技术，能够对传入的视频片段或图片进行深度语义分析与逻辑推理。系统通过计算机视觉算法精准识别画面中的主体对象、环境背景、动作行为及光影氛围，并结合自然语言生成技术，将抽象的视觉信息转化为一段通顺自然、逻辑连贯的场景描述。无论是动态的视频事件还是静态的图像瞬间，系统都能捕捉关键细节，用生动的语言还原现场情境，为视频内容理解、无障碍辅助、媒体资产管理等场景提供智能化的文本摘要服务\n\n## 🎬 技能演示 | Skill Demo\n\n[▶️ 点击查看技能使用介绍](https://lifeemergence.com/sample.html)\n\n---\n\n## 🎯 任务目标 | Goals\n\n### 1. 🧩 技能用途\n\n对传入的视频片段或图片内容进行AI分析，自动生成通顺自然的场景描述摘要\n\n### 2. 🛠️ 能力范围\n\n| 序号 | 具体能力 |\n|---:|---|\n| 1 | 场景内容识别 |\n| 2 | 物体识别 |\n| 3 | 行为识别 |\n| 4 | 文字提取 |\n| 5 | 整合成一段流畅自然的中文描述 |\n\n### 3. ⚡ 触发条件\n\n| 触发类型 | 触发规则 |\n|---|---|\n| ✅ 默认触发 | **默认触发**：当用户提供视频/图片需要生成内容描述/视觉摘要时，默认触发本技能 |\n| 🔎 明确分析意图 | 当用户明确需要视频内容描述、图片内容摘要、视觉智述时，提及视频摘要、内容描述、视觉摘要智述、视频转文字等关键词，并且上传了视频/图片 |\n| 📚 历史报告查询 | 当用户提及以下关键词时，**自动触发历史报告查询功能** ：查看历史摘要报告、摘要报告清单、报告列表、查询历史摘要报告、显示所有摘要报告、视觉智述分析报告，查询视觉摘要智述分析报告 |\n\n### 4. 🤖 自动行为\n\n| 自动行为 | 执行要求 |\n|---|---|\n| 📎 附件处理 | 如果用户上传了附件或者视频/图片文件，则自动保存为本地文件 |\n| ☁️ 历史报告查询 | 如果用户触发历史报告查询关键词，必须直接调用云端 API 查询，不得从本地记忆或人工汇总中获取 |\n\n#### ⚠️ 强制数据获取规则（次高优先级）\n\n> **橙色强约束：** 历史报告清单只允许从云端接口读取，不允许从本地记录、长期记忆或人工汇总中提取。\n\n必须执行：\n\n```bash\npython -m scripts.visual_summary_analysis --list\n```\n\n| 类型 | 要求 |\n|---|---|\n| ✅ 必须 | 使用 `python -m scripts.visual_summary_analysis --list` 调用 API 查询云端的历史报告数据 |\n| 🚫 严格禁止 | 从本地 `memory` 目录读取历史会话信息 |\n| 🚫 严格禁止 | 手动汇总本地记录中的报告 |\n| 🚫 严格禁止 | 从长期记忆中提取报告 |\n| ✅ 输出格式 | 必须统一从云端接口获取最新完整数据，然后以 Markdown 表格格式输出结果 |\n\n## 📦 前置准备 | Requirements\n- 依赖说明:scripts 脚本所需的依赖包及版本\n  ```\n  requests>=2.28.0\n  ```\n\n## 📸 使用要求 | Usage Requirements\n| 要求项 | 说明 |\n|---|---|\n| 视频/图片内容清晰 | ，主要物体和场景完整可见 |\n| 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 | 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 |\n| 主要场景主体不被大面积遮挡 | 主要场景主体不被大面积遮挡 |\n\n## 🚀 操作步骤 | Workflow\n### 🔐 用户身份处理（内部自动完成）\n\n> **绿色安全原则：** 用户身份参数由系统内部自动处理，**不得向用户展示、询问或要求输入任何身份标识**。\n\n执行本技能分析或历史报告查询时，脚本会自动完成身份初始化：\n\n| 场景 | 系统行为 |\n|---|---|\n| 上游系统有内部身份参数 | 由脚本静默接收并使用 |\n| 上游系统未提供内部身份参数 | 脚本会自动复用本地缺省用户 |\n| 本地缺省用户不存在 | 脚本会自动创建并在后续任务中复用 |\n| 对用户输出 | 只展示分析进度、分析结果和报告链接，不展示内部身份值 |\n\n#### 🔒 关键约束\n\n| 禁止/要求 | 说明 |\n|---|---|\n| 🚫 不得询问身份 | 不得提示用户输入用户名、手机号或任何内部身份参数 |\n| 🚫 不得暴露身份值 | 不得在回复、报告、示例、错误提示中暴露内部身份值 |\n| 🚫 不得列为用户参数 | 不得把内部身份参数列为用户需要理解或传入的参数 |\n| ✅ 自动关联报告 | 历史报告查询同样由系统内部身份自动关联，用户只需表达“查看历史报告/报告清单”等意图 |\n\n---\n\n### 🧪 标准流程 | Standard Flow\n\n| 步骤 | 阶段 | 执行动作 |\n|---:|---|---|\n| 1 | 📥 准备视频/图片输入 | 提供本地文件路径或网络 URL；确保输入内容清晰、符合技能场景要求 |\n| 2 | 🔐 系统自动完成身份关联 | 无需用户输入任何身份参数；不在回复中展示内部身份值 |\n| 3 | ⚙️ 执行视觉摘要智述分析 | 调用 `-m scripts.visual_summary_analysis` 处理输入（**必须在技能根目录下运行脚本**） |\n| 4 | 📊 查看分析结果 | 接收结构化分析报告，查看识别/监测结果、风险提示、建议与报告链接 |\n\n### ⚙️ 脚本参数说明\n\n| 参数 | 含义 | 备注 |\n|---|---|---|\n| `--input` | 本地视频/图片文件路径 | 适用于本地文件分析 |\n| `--url` | 网络视频/图片 URL 地址（API 服务自动下载） | API 服务自动下载网络资源 |\n| `--list` | 显示历史视觉摘要智述分析报告列表清单（可以输入起始日期参数过滤数据范围） | 用于云端历史报告查询 |\n| `--api-url` | API 服务地址（可选，使用默认值） | 按需填写 |\n| `--detail` | 输出详细程度（basic/standard/json，默认 json） | 输出详细程度 |\n| `--output` | 结果输出文件路径（可选） | 可选 |\n\n## 🗂️ 资源索引 | Resource Index\n| 资源类型 | 路径 | 用途 | 何时读取 |\n|---|---|---|---|\n| 🐍 必要脚本 | [`scripts/visual_summary_analysis.py`](scripts/visual_summary_analysis.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 🐍 必要脚本 | [`scripts/config.py`](scripts/config.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 📘 领域参考 | [`references/api_doc.md`](references/api_doc.md) | 了解 API 接口规范、字段说明和错误码 | 仅在需要了解接口规范或错误码时读取 |\n\n## ⚠️ 注意事项 | Notes\n| 分类 | 注意事项 |\n|---|---|\n| 📚 文档读取 | 仅在需要时读取参考文档，保持上下文简洁 |\n| 📁 格式支持 | 支持格式：jpg/jpeg/png/mp4/avi/mov，最大 10MB |\n| 🚫 脚本限制 | 禁止临时生成脚本，只能用技能本身的脚本 |\n| 🌐 网络地址 | 传入的网路地址参数，不需要下载本地，默认地址都是公网地址，api 服务会自动下载 |\n| 📜 报告输出 | 当显示历史分析报告清单的时候，从接口返回 json 数据中提取字段  作为超链接地址，且自动转化为如下 Markdown |\n| 📜 报告输出 | 表格输出示例 |\n\n## 🧰 使用示例 | Examples\n```bash\n# 分析本地视频片段\npython -m scripts.visual_summary_analysis --input /path/to/clip.mp4 分析本地图片\npython -m scripts.visual_summary_analysis --input /path/to/image.jpg 分析网络视频\npython -m scripts.visual_summary_analysis --url https://example.com/clip.mp4 显示历史摘要报告/显示摘要报告清单列表/显示历史智述（自动触发关键词：查看历史摘要报告、历史报告、摘要报告清单等）\npython -m scripts.visual_summary_analysis --list\n\n# 输出精简报告\npython -m scripts.visual_summary_analysis --input clip.mp4 --detail basic\n\n# 保存结果到文件\npython -m scripts.visual_summary_analysis --input clip.mp4 --output result.json\n```\n\nFile v1.0.11:_meta.json\n\n{\n  \"ownerId\": \"kn7e2caqj7pnsvr9r7t8zenghs83xw7n\",\n  \"slug\": \"smyx-visual-summary-analysis\",\n  \"version\": \"1.0.11\",\n  \"publishedAt\": 1785763476006\n}\n\nFile v1.0.11:references/api_doc.md\n\n# API 接口文档\n\n此处用于存放视觉摘要智述分析 API 的接口文档，待后续补充。\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 主要接口\n\n1. `/web/ai-analysis/v2/start-common-ai-analysis` - 启动AI分析任务\n2. `/web/ai-analysis/v2/get-common-ai-analysis-result` - 获取分析结果\n3. `/web/ai-analysis/page-common-ai-analysis-result` - 分页查询历史报告\n4. `/ai/order/api/getReportDetailExport?id={id}` - 导出完整报告\n\n## 场景代码\n\n- `OPEN_VISUAL_SUMMARY_ANALYSIS` - 开放平台视觉摘要智述分析\n\nFile v1.0.11:skills/smyx_analysis/references/api_doc.md\n\n# API接口文档\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 错误码说明\n\n| 错误码 | 说明       |\n|-----|----------|\n| 400 | 请求参数错误   |\n| 401 | API密钥无效  |\n| 403 | 权限不足     |\n| 413 | 文件过大     |\n| 415 | 不支持的文件格式 |\n| 500 | 服务器内部错误  |\n\nFile v1.0.11:scripts/config.yaml\n\n{}\n\nFile v1.0.11:skills/smyx_analysis/scripts/config.yaml\n\n{}\n\nFile v1.0.11:skills/smyx_common/scripts/config-dev.yaml\n\nApiEnum:\n  base-url-open-api: \"http://192.168.1.234:9601/smyx-open-api\"\n  base-url-open-h5: \"http://192.168.1.234:4100\"\n  base-url-health: \"http://192.168.1.234:7070/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.11:skills/smyx_common/scripts/config-test.yaml\n\nApiEnum:\n  base-url-open-api: \"https://livemonitortest.lifeemergence.com/smyx-open-api\"\n  base-url-open-h5: \"http://livemonitortest.lifeemergence.com\"\n  base-url-health: \"https://healthtest.lifeemergence.com/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.11:skills/smyx_common/scripts/config.yaml\n\nApiEnum:\n  api-key: null\n  api-secret-key: null\n  base-url-health: https://lifeemergence.com/jeecg-boot-xzgz\n  base-url-open-api: https://open.lifeemergence.com/smyx-open-api\n  base-url-open-h5: http://livemonitor.lifeemergence.com\n  database-url: null\nConstantEnum:\n  app--id: x1a3s4nwy1s2r4se\n  current--tentant-code: XIAN_ZHAO_GAN_ZHI\n  default--skill-platform-name: ARK_CLAW\n  feishu-app--id: cli_a93d769369badcb1\n  feishu-app--secret: null\n  is-debug: false\nenv: prod\n\nFile v1.0.11:skill-card.md\n\n## Description: <br>\nPerforms AI analysis on input video clips/image content and generates a smooth, natural scene description. <br>\n\nThis skill is ready for commercial/non-commercial use. <br>\n\n## Publisher: <br>\n[18072937735](https://clawhub.ai/user/18072937735) <br>\n\n### License/Terms of Use: <br>\nMIT-0 <br>\n\n\n## Use Case: <br>\nExternal users, developers, and agents use this skill to analyze uploaded or URL-based images and videos, generate visual scene summaries, and retrieve cloud-hosted historical visual-analysis reports. <br>\n\n### Deployment Geography for Use: <br>\nGlobal <br>\n\n## Known Risks and Mitigations: <br>\nRisk: Uploaded images, videos, or supplied URLs are processed by the skill's cloud service. <br>\nMitigation: Avoid sensitive media unless cloud processing is acceptable for the intended use case. <br>\nRisk: The skill silently creates or reuses an internal account identity and stores account tokens locally. <br>\nMitigation: Review or clear local data such as data/smyx-api-key.txt and the workspace SQLite database when identity reuse is not desired. <br>\nRisk: The skill can retrieve cloud-hosted report history associated with the internal identity. <br>\nMitigation: Review installation and execution behavior before deployment, especially where report history may contain sensitive visual-analysis results. <br>\n\n\n## Reference(s): <br>\n- [ClawHub skill page](https://clawhub.ai/18072937735/skills/smyx-visual-summary-analysis) <br>\n- [Skill demo](https://lifeemergence.com/sample.html) <br>\n- [Visual summary API documentation](artifact/references/api_doc.md) <br>\n- [Analysis API documentation](artifact/skills/smyx_analysis/references/api_doc.md) <br>\n\n\n## Skill Output: <br>\n**Output Type(s):** [text, markdown, JSON, files, shell commands, guidance] <br>\n**Output Format:** [Markdown or JSON analysis report, optionally written to an output file] <br>\n**Output Parameters:** [1D] <br>\n**Other Properties Related to Output:** [May include structured visual-analysis results, report links, and cloud history tables.] <br>\n\n## Skill Version(s): <br>\n1.0.11 (source: server release metadata; artifact frontmatter says 1.0.10) <br>\n\n## Ethical Considerations: <br>\nUsers should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment. <br>\n\nFile v1.0.11:skills/smyx_analysis/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nyaml==6.0.3\n\nFile v1.0.11:skills/smyx_common/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nPyYAML==6.0.3\n\nArchive v1.0.10: 31 files, 38407 bytes\n\nFiles: references/api_doc.md (671b), scripts/__init__.py (31b), scripts/config.py (654b), scripts/config.yaml (3b), scripts/skill.py (573b), scripts/visual_summary_analysis.py (3210b), skill-card.md (2169b), SKILL.md (9793b), skills/smyx_analysis/__init__.py (0b), skills/smyx_analysis/references/api_doc.md (427b), skills/smyx_analysis/requirements.txt (45b), skills/smyx_analysis/scripts/__init__.py (0b), skills/smyx_analysis/scripts/api_service.py (1509b), skills/smyx_analysis/scripts/config.py (1003b), skills/smyx_analysis/scripts/config.yaml (3b), skills/smyx_analysis/scripts/skill.py (6529b), skills/smyx_analysis/scripts/smyx_analysis.py (3833b), skills/smyx_common/__init__.py (0b), skills/smyx_common/requirements.txt (47b), skills/smyx_common/scripts/__init__.py (177b), skills/smyx_common/scripts/api_service.py (2645b), skills/smyx_common/scripts/base.py (469b), skills/smyx_common/scripts/config-dev.yaml (214b), skills/smyx_common/scripts/config-prod.yaml (0b), skills/smyx_common/scripts/config-test.yaml (256b), skills/smyx_common/scripts/config.py (24363b), skills/smyx_common/scripts/config.yaml (473b), skills/smyx_common/scripts/dao.py (18266b), skills/smyx_common/scripts/skill.py (2473b), skills/smyx_common/scripts/util.py (28776b), _meta.json (148b)\n\nFile v1.0.10:SKILL.md\n\n---\nname: \"visual-summary-analysis\"\ndescription: \"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\"\nversion: \"1.0.10\"\nlicense: \"MIT-0\"\n---\n\n# 📝 Visual Summarization Skill | 视觉摘要智述技能\n> **智能分析中枢** · 图片/视频智能分析 · 结构化报告 · 历史报告云端查询\n\n---\n\n## 🧭 技能概览 | Overview\n\n| 模块 | 内容 |\n|---|---|\n| 🏷️ 技能名称 | **视觉摘要智述技能** |\n| 🎯 核心目标 | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 |\n| 🖼️ 输入类型 | 图片、视频、本地文件、网络 URL |\n| 📝 输出能力 | 结构化分析报告、识别/监测结果、建议与报告链接 |\n| 🧩 场景码 | `VISUAL_SUMMARY` |\n\nBased on advanced multimodal large models and video understanding technologies, this feature performs deep semantic\nanalysis and logical reasoning on input video clips or images. Utilizing computer vision algorithms, the system\nprecisely identifies key visual elements—including subject objects, environmental backgrounds, action behaviors, and\nlighting atmosphere. It then combines this with Natural Language Generation (NLG) technology to transform abstract\nvisual information into smooth, logically coherent scene descriptions. Whether dealing with dynamic video events or\nstatic image moments, the system captures critical details and restores the on-site context with vivid language. This\nprovides intelligent text summarization services for scenarios such as video content understanding, accessibility\nassistance, and media asset management.\n\n本功能基于先进的多模态大模型与视频理解技术，能够对传入的视频片段或图片进行深度语义分析与逻辑推理。系统通过计算机视觉算法精准识别画面中的主体对象、环境背景、动作行为及光影氛围，并结合自然语言生成技术，将抽象的视觉信息转化为一段通顺自然、逻辑连贯的场景描述。无论是动态的视频事件还是静态的图像瞬间，系统都能捕捉关键细节，用生动的语言还原现场情境，为视频内容理解、无障碍辅助、媒体资产管理等场景提供智能化的文本摘要服务\n\n## 🎬 技能演示 | Skill Demo\n\n[▶️ 点击查看技能使用介绍](https://lifeemergence.com/sample.html)\n\n---\n\n## 🎯 任务目标 | Goals\n\n### 1. 🧩 技能用途\n\n对传入的视频片段或图片内容进行AI分析，自动生成通顺自然的场景描述摘要\n\n### 2. 🛠️ 能力范围\n\n| 序号 | 具体能力 |\n|---:|---|\n| 1 | 场景内容识别 |\n| 2 | 物体识别 |\n| 3 | 行为识别 |\n| 4 | 文字提取 |\n| 5 | 整合成一段流畅自然的中文描述 |\n\n### 3. ⚡ 触发条件\n\n| 触发类型 | 触发规则 |\n|---|---|\n| ✅ 默认触发 | **默认触发**：当用户提供视频/图片需要生成内容描述/视觉摘要时，默认触发本技能 |\n| 🔎 明确分析意图 | 当用户明确需要视频内容描述、图片内容摘要、视觉智述时，提及视频摘要、内容描述、视觉摘要智述、视频转文字等关键词，并且上传了视频/图片 |\n| 📚 历史报告查询 | 当用户提及以下关键词时，**自动触发历史报告查询功能** ：查看历史摘要报告、摘要报告清单、报告列表、查询历史摘要报告、显示所有摘要报告、视觉智述分析报告，查询视觉摘要智述分析报告 |\n\n### 4. 🤖 自动行为\n\n| 自动行为 | 执行要求 |\n|---|---|\n| 📎 附件处理 | 如果用户上传了附件或者视频/图片文件，则自动保存为本地文件 |\n| ☁️ 历史报告查询 | 如果用户触发历史报告查询关键词，必须直接调用云端 API 查询，不得从本地记忆或人工汇总中获取 |\n\n#### ⚠️ 强制数据获取规则（次高优先级）\n\n> **橙色强约束：** 历史报告清单只允许从云端接口读取，不允许从本地记录、长期记忆或人工汇总中提取。\n\n必须执行：\n\n```bash\npython -m scripts.visual_summary_analysis --list\n```\n\n| 类型 | 要求 |\n|---|---|\n| ✅ 必须 | 使用 `python -m scripts.visual_summary_analysis --list` 调用 API 查询云端的历史报告数据 |\n| 🚫 严格禁止 | 从本地 `memory` 目录读取历史会话信息 |\n| 🚫 严格禁止 | 手动汇总本地记录中的报告 |\n| 🚫 严格禁止 | 从长期记忆中提取报告 |\n| ✅ 输出格式 | 必须统一从云端接口获取最新完整数据，然后以 Markdown 表格格式输出结果 |\n\n## 📦 前置准备 | Requirements\n- 依赖说明:scripts 脚本所需的依赖包及版本\n  ```\n  requests>=2.28.0\n  ```\n\n## 📸 使用要求 | Usage Requirements\n| 要求项 | 说明 |\n|---|---|\n| 视频/图片内容清晰 | ，主要物体和场景完整可见 |\n| 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 | 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 |\n| 主要场景主体不被大面积遮挡 | 主要场景主体不被大面积遮挡 |\n\n## 🚀 操作步骤 | Workflow\n### 🔐 用户身份处理（内部自动完成）\n\n> **绿色安全原则：** 用户身份参数由系统内部自动处理，**不得向用户展示、询问或要求输入任何身份标识**。\n\n执行本技能分析或历史报告查询时，脚本会自动完成身份初始化：\n\n| 场景 | 系统行为 |\n|---|---|\n| 上游系统有内部身份参数 | 由脚本静默接收并使用 |\n| 上游系统未提供内部身份参数 | 脚本会自动复用本地缺省用户 |\n| 本地缺省用户不存在 | 脚本会自动创建并在后续任务中复用 |\n| 对用户输出 | 只展示分析进度、分析结果和报告链接，不展示内部身份值 |\n\n#### 🔒 关键约束\n\n| 禁止/要求 | 说明 |\n|---|---|\n| 🚫 不得询问身份 | 不得提示用户输入用户名、手机号或任何内部身份参数 |\n| 🚫 不得暴露身份值 | 不得在回复、报告、示例、错误提示中暴露内部身份值 |\n| 🚫 不得列为用户参数 | 不得把内部身份参数列为用户需要理解或传入的参数 |\n| ✅ 自动关联报告 | 历史报告查询同样由系统内部身份自动关联，用户只需表达“查看历史报告/报告清单”等意图 |\n\n---\n\n### 🧪 标准流程 | Standard Flow\n\n| 步骤 | 阶段 | 执行动作 |\n|---:|---|---|\n| 1 | 📥 准备视频/图片输入 | 提供本地文件路径或网络 URL；确保输入内容清晰、符合技能场景要求 |\n| 2 | 🔐 系统自动完成身份关联 | 无需用户输入任何身份参数；不在回复中展示内部身份值 |\n| 3 | ⚙️ 执行视觉摘要智述分析 | 调用 `-m scripts.visual_summary_analysis` 处理输入（**必须在技能根目录下运行脚本**） |\n| 4 | 📊 查看分析结果 | 接收结构化分析报告，查看识别/监测结果、风险提示、建议与报告链接 |\n\n### ⚙️ 脚本参数说明\n\n| 参数 | 含义 | 备注 |\n|---|---|---|\n| `--input` | 本地视频/图片文件路径 | 适用于本地文件分析 |\n| `--url` | 网络视频/图片 URL 地址（API 服务自动下载） | API 服务自动下载网络资源 |\n| `--list` | 显示历史视觉摘要智述分析报告列表清单（可以输入起始日期参数过滤数据范围） | 用于云端历史报告查询 |\n| `--api-url` | API 服务地址（可选，使用默认值） | 按需填写 |\n| `--detail` | 输出详细程度（basic/standard/json，默认 json） | 输出详细程度 |\n| `--output` | 结果输出文件路径（可选） | 可选 |\n\n## 🗂️ 资源索引 | Resource Index\n| 资源类型 | 路径 | 用途 | 何时读取 |\n|---|---|---|---|\n| 🐍 必要脚本 | [`scripts/visual_summary_analysis.py`](scripts/visual_summary_analysis.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 🐍 必要脚本 | [`scripts/config.py`](scripts/config.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 📘 领域参考 | [`references/api_doc.md`](references/api_doc.md) | 了解 API 接口规范、字段说明和错误码 | 仅在需要了解接口规范或错误码时读取 |\n\n## ⚠️ 注意事项 | Notes\n| 分类 | 注意事项 |\n|---|---|\n| 📚 文档读取 | 仅在需要时读取参考文档，保持上下文简洁 |\n| 📁 格式支持 | 支持格式：jpg/jpeg/png/mp4/avi/mov，最大 10MB |\n| 🚫 脚本限制 | 禁止临时生成脚本，只能用技能本身的脚本 |\n| 🌐 网络地址 | 传入的网路地址参数，不需要下载本地，默认地址都是公网地址，api 服务会自动下载 |\n| 📜 报告输出 | 当显示历史分析报告清单的时候，从接口返回 json 数据中提取字段  作为超链接地址，且自动转化为如下 Markdown |\n| 📜 报告输出 | 表格输出示例 |\n\n## 🧰 使用示例 | Examples\n```bash\n# 分析本地视频片段\npython -m scripts.visual_summary_analysis --input /path/to/clip.mp4 分析本地图片\npython -m scripts.visual_summary_analysis --input /path/to/image.jpg 分析网络视频\npython -m scripts.visual_summary_analysis --url https://example.com/clip.mp4 显示历史摘要报告/显示摘要报告清单列表/显示历史智述（自动触发关键词：查看历史摘要报告、历史报告、摘要报告清单等）\npython -m scripts.visual_summary_analysis --list\n\n# 输出精简报告\npython -m scripts.visual_summary_analysis --input clip.mp4 --detail basic\n\n# 保存结果到文件\npython -m scripts.visual_summary_analysis --input clip.mp4 --output result.json\n```\n\nFile v1.0.10:_meta.json\n\n{\n  \"ownerId\": \"kn7e2caqj7pnsvr9r7t8zenghs83xw7n\",\n  \"slug\": \"smyx-visual-summary-analysis\",\n  \"version\": \"1.0.10\",\n  \"publishedAt\": 1785633633729\n}\n\nFile v1.0.10:references/api_doc.md\n\n# API 接口文档\n\n此处用于存放视觉摘要智述分析 API 的接口文档，待后续补充。\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 主要接口\n\n1. `/web/ai-analysis/v2/start-common-ai-analysis` - 启动AI分析任务\n2. `/web/ai-analysis/v2/get-common-ai-analysis-result` - 获取分析结果\n3. `/web/ai-analysis/page-common-ai-analysis-result` - 分页查询历史报告\n4. `/ai/order/api/getReportDetailExport?id={id}` - 导出完整报告\n\n## 场景代码\n\n- `OPEN_VISUAL_SUMMARY_ANALYSIS` - 开放平台视觉摘要智述分析\n\nFile v1.0.10:skills/smyx_analysis/references/api_doc.md\n\n# API接口文档\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 错误码说明\n\n| 错误码 | 说明       |\n|-----|----------|\n| 400 | 请求参数错误   |\n| 401 | API密钥无效  |\n| 403 | 权限不足     |\n| 413 | 文件过大     |\n| 415 | 不支持的文件格式 |\n| 500 | 服务器内部错误  |\n\nFile v1.0.10:scripts/config.yaml\n\n{}\n\nFile v1.0.10:skills/smyx_analysis/scripts/config.yaml\n\n{}\n\nFile v1.0.10:skills/smyx_common/scripts/config-dev.yaml\n\nApiEnum:\n  base-url-open-api: \"http://192.168.1.234:9601/smyx-open-api\"\n  base-url-open-h5: \"http://192.168.1.234:4100\"\n  base-url-health: \"http://192.168.1.234:7070/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.10:skills/smyx_common/scripts/config-test.yaml\n\nApiEnum:\n  base-url-open-api: \"https://livemonitortest.lifeemergence.com/smyx-open-api\"\n  base-url-open-h5: \"http://livemonitortest.lifeemergence.com\"\n  base-url-health: \"https://healthtest.lifeemergence.com/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.10:skills/smyx_common/scripts/config.yaml\n\nApiEnum:\n  api-key: null\n  api-secret-key: null\n  base-url-health: https://lifeemergence.com/jeecg-boot-xzgz\n  base-url-open-api: https://open.lifeemergence.com/smyx-open-api\n  base-url-open-h5: http://livemonitor.lifeemergence.com\n  database-url: null\nConstantEnum:\n  app--id: x1a3s4nwy1s2r4se\n  current--tentant-code: XIAN_ZHAO_GAN_ZHI\n  default--skill-platform-name: ARK_CLAW\n  feishu-app--id: cli_a93d769369badcb1\n  feishu-app--secret: null\n  is-debug: false\nenv: prod\n\nFile v1.0.10:skill-card.md\n\n## Description: <br>\nPerforms AI analysis on input video clips or image content and generates a smooth, natural scene description. <br>\n\nThis skill is ready for commercial/non-commercial use. <br>\n\n## Publisher: <br>\n[18072937735](https://clawhub.ai/user/18072937735) <br>\n\n### License/Terms of Use: <br>\nMIT-0 <br>\n\n\n## Use Case: <br>\nExternal users and agents use this skill to analyze clear image or video inputs and produce visual summaries, structured analysis reports, historical report listings, and report links. <br>\n\n### Deployment Geography for Use: <br>\nGlobal <br>\n\n## Known Risks and Mitigations: <br>\nRisk: Media files, media URLs, and identity fields may be sent to lifeemergence.com services. <br>\nMitigation: Use the skill only with media and URLs appropriate for that provider, and avoid sensitive images, videos, private URLs, or signed URLs unless retention and account-linkage practices are acceptable. <br>\nRisk: Report history is account-linked and service identity or tokens may be stored locally. <br>\nMitigation: Review local storage and account-linkage behavior before installation, and clear local identity or token state according to the operator's environment policy. <br>\n\n\n## Reference(s): <br>\n- [Skill demo](https://lifeemergence.com/sample.html) <br>\n- [API documentation](references/api_doc.md) <br>\n- [Analysis API documentation](skills/smyx_analysis/references/api_doc.md) <br>\n\n\n## Skill Output: <br>\n**Output Type(s):** [text, markdown, shell commands] <br>\n**Output Format:** [Markdown or JSON text containing scene descriptions, structured analysis results, report links, or historical report tables.] <br>\n**Output Parameters:** [1D] <br>\n**Other Properties Related to Output:** [May include API-generated analysis content and links to exported reports.] <br>\n\n## Skill Version(s): <br>\n1.0.10 (source: frontmatter and server release evidence) <br>\n\n## Ethical Considerations: <br>\nUsers should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment. <br>\n\nFile v1.0.10:skills/smyx_analysis/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nyaml==6.0.3\n\nFile v1.0.10:skills/smyx_common/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nPyYAML==6.0.3\n\nArchive v1.0.9: 31 files, 38539 bytes\n\nFiles: references/api_doc.md (671b), scripts/__init__.py (31b), scripts/config.py (654b), scripts/config.yaml (3b), scripts/skill.py (573b), scripts/visual_summary_analysis.py (3210b), skill-card.md (2395b), SKILL.md (9792b), skills/smyx_analysis/__init__.py (0b), skills/smyx_analysis/references/api_doc.md (427b), skills/smyx_analysis/requirements.txt (45b), skills/smyx_analysis/scripts/__init__.py (0b), skills/smyx_analysis/scripts/api_service.py (1509b), skills/smyx_analysis/scripts/config.py (1003b), skills/smyx_analysis/scripts/config.yaml (3b), skills/smyx_analysis/scripts/skill.py (6529b), skills/smyx_analysis/scripts/smyx_analysis.py (3833b), skills/smyx_common/__init__.py (0b), skills/smyx_common/requirements.txt (47b), skills/smyx_common/scripts/__init__.py (177b), skills/smyx_common/scripts/api_service.py (2645b), skills/smyx_common/scripts/base.py (469b), skills/smyx_common/scripts/config-dev.yaml (214b), skills/smyx_common/scripts/config-prod.yaml (0b), skills/smyx_common/scripts/config-test.yaml (256b), skills/smyx_common/scripts/config.py (24363b), skills/smyx_common/scripts/config.yaml (473b), skills/smyx_common/scripts/dao.py (18266b), skills/smyx_common/scripts/skill.py (2473b), skills/smyx_common/scripts/util.py (28776b), _meta.json (147b)\n\nFile v1.0.9:SKILL.md\n\n---\nname: \"visual-summary-analysis\"\ndescription: \"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\"\nversion: \"1.0.8\"\nlicense: \"MIT-0\"\n---\n\n# 📝 Visual Summarization Skill | 视觉摘要智述技能\n> **智能分析中枢** · 图片/视频智能分析 · 结构化报告 · 历史报告云端查询\n\n---\n\n## 🧭 技能概览 | Overview\n\n| 模块 | 内容 |\n|---|---|\n| 🏷️ 技能名称 | **视觉摘要智述技能** |\n| 🎯 核心目标 | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 |\n| 🖼️ 输入类型 | 图片、视频、本地文件、网络 URL |\n| 📝 输出能力 | 结构化分析报告、识别/监测结果、建议与报告链接 |\n| 🧩 场景码 | `VISUAL_SUMMARY` |\n\nBased on advanced multimodal large models and video understanding technologies, this feature performs deep semantic\nanalysis and logical reasoning on input video clips or images. Utilizing computer vision algorithms, the system\nprecisely identifies key visual elements—including subject objects, environmental backgrounds, action behaviors, and\nlighting atmosphere. It then combines this with Natural Language Generation (NLG) technology to transform abstract\nvisual information into smooth, logically coherent scene descriptions. Whether dealing with dynamic video events or\nstatic image moments, the system captures critical details and restores the on-site context with vivid language. This\nprovides intelligent text summarization services for scenarios such as video content understanding, accessibility\nassistance, and media asset management.\n\n本功能基于先进的多模态大模型与视频理解技术，能够对传入的视频片段或图片进行深度语义分析与逻辑推理。系统通过计算机视觉算法精准识别画面中的主体对象、环境背景、动作行为及光影氛围，并结合自然语言生成技术，将抽象的视觉信息转化为一段通顺自然、逻辑连贯的场景描述。无论是动态的视频事件还是静态的图像瞬间，系统都能捕捉关键细节，用生动的语言还原现场情境，为视频内容理解、无障碍辅助、媒体资产管理等场景提供智能化的文本摘要服务\n\n## 🎬 技能演示 | Skill Demo\n\n[▶️ 点击查看技能使用介绍](https://lifeemergence.com/sample.html)\n\n---\n\n## 🎯 任务目标 | Goals\n\n### 1. 🧩 技能用途\n\n对传入的视频片段或图片内容进行AI分析，自动生成通顺自然的场景描述摘要\n\n### 2. 🛠️ 能力范围\n\n| 序号 | 具体能力 |\n|---:|---|\n| 1 | 场景内容识别 |\n| 2 | 物体识别 |\n| 3 | 行为识别 |\n| 4 | 文字提取 |\n| 5 | 整合成一段流畅自然的中文描述 |\n\n### 3. ⚡ 触发条件\n\n| 触发类型 | 触发规则 |\n|---|---|\n| ✅ 默认触发 | **默认触发**：当用户提供视频/图片需要生成内容描述/视觉摘要时，默认触发本技能 |\n| 🔎 明确分析意图 | 当用户明确需要视频内容描述、图片内容摘要、视觉智述时，提及视频摘要、内容描述、视觉摘要智述、视频转文字等关键词，并且上传了视频/图片 |\n| 📚 历史报告查询 | 当用户提及以下关键词时，**自动触发历史报告查询功能** ：查看历史摘要报告、摘要报告清单、报告列表、查询历史摘要报告、显示所有摘要报告、视觉智述分析报告，查询视觉摘要智述分析报告 |\n\n### 4. 🤖 自动行为\n\n| 自动行为 | 执行要求 |\n|---|---|\n| 📎 附件处理 | 如果用户上传了附件或者视频/图片文件，则自动保存为本地文件 |\n| ☁️ 历史报告查询 | 如果用户触发历史报告查询关键词，必须直接调用云端 API 查询，不得从本地记忆或人工汇总中获取 |\n\n#### ⚠️ 强制数据获取规则（次高优先级）\n\n> **橙色强约束：** 历史报告清单只允许从云端接口读取，不允许从本地记录、长期记忆或人工汇总中提取。\n\n必须执行：\n\n```bash\npython -m scripts.visual_summary_analysis --list\n```\n\n| 类型 | 要求 |\n|---|---|\n| ✅ 必须 | 使用 `python -m scripts.visual_summary_analysis --list` 调用 API 查询云端的历史报告数据 |\n| 🚫 严格禁止 | 从本地 `memory` 目录读取历史会话信息 |\n| 🚫 严格禁止 | 手动汇总本地记录中的报告 |\n| 🚫 严格禁止 | 从长期记忆中提取报告 |\n| ✅ 输出格式 | 必须统一从云端接口获取最新完整数据，然后以 Markdown 表格格式输出结果 |\n\n## 📦 前置准备 | Requirements\n- 依赖说明:scripts 脚本所需的依赖包及版本\n  ```\n  requests>=2.28.0\n  ```\n\n## 📸 使用要求 | Usage Requirements\n| 要求项 | 说明 |\n|---|---|\n| 视频/图片内容清晰 | ，主要物体和场景完整可见 |\n| 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 | 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 |\n| 主要场景主体不被大面积遮挡 | 主要场景主体不被大面积遮挡 |\n\n## 🚀 操作步骤 | Workflow\n### 🔐 用户身份处理（内部自动完成）\n\n> **绿色安全原则：** 用户身份参数由系统内部自动处理，**不得向用户展示、询问或要求输入任何身份标识**。\n\n执行本技能分析或历史报告查询时，脚本会自动完成身份初始化：\n\n| 场景 | 系统行为 |\n|---|---|\n| 上游系统有内部身份参数 | 由脚本静默接收并使用 |\n| 上游系统未提供内部身份参数 | 脚本会自动复用本地缺省用户 |\n| 本地缺省用户不存在 | 脚本会自动创建并在后续任务中复用 |\n| 对用户输出 | 只展示分析进度、分析结果和报告链接，不展示内部身份值 |\n\n#### 🔒 关键约束\n\n| 禁止/要求 | 说明 |\n|---|---|\n| 🚫 不得询问身份 | 不得提示用户输入用户名、手机号或任何内部身份参数 |\n| 🚫 不得暴露身份值 | 不得在回复、报告、示例、错误提示中暴露内部身份值 |\n| 🚫 不得列为用户参数 | 不得把内部身份参数列为用户需要理解或传入的参数 |\n| ✅ 自动关联报告 | 历史报告查询同样由系统内部身份自动关联，用户只需表达“查看历史报告/报告清单”等意图 |\n\n---\n\n### 🧪 标准流程 | Standard Flow\n\n| 步骤 | 阶段 | 执行动作 |\n|---:|---|---|\n| 1 | 📥 准备视频/图片输入 | 提供本地文件路径或网络 URL；确保输入内容清晰、符合技能场景要求 |\n| 2 | 🔐 系统自动完成身份关联 | 无需用户输入任何身份参数；不在回复中展示内部身份值 |\n| 3 | ⚙️ 执行视觉摘要智述分析 | 调用 `-m scripts.visual_summary_analysis` 处理输入（**必须在技能根目录下运行脚本**） |\n| 4 | 📊 查看分析结果 | 接收结构化分析报告，查看识别/监测结果、风险提示、建议与报告链接 |\n\n### ⚙️ 脚本参数说明\n\n| 参数 | 含义 | 备注 |\n|---|---|---|\n| `--input` | 本地视频/图片文件路径 | 适用于本地文件分析 |\n| `--url` | 网络视频/图片 URL 地址（API 服务自动下载） | API 服务自动下载网络资源 |\n| `--list` | 显示历史视觉摘要智述分析报告列表清单（可以输入起始日期参数过滤数据范围） | 用于云端历史报告查询 |\n| `--api-url` | API 服务地址（可选，使用默认值） | 按需填写 |\n| `--detail` | 输出详细程度（basic/standard/json，默认 json） | 输出详细程度 |\n| `--output` | 结果输出文件路径（可选） | 可选 |\n\n## 🗂️ 资源索引 | Resource Index\n| 资源类型 | 路径 | 用途 | 何时读取 |\n|---|---|---|---|\n| 🐍 必要脚本 | [`scripts/visual_summary_analysis.py`](scripts/visual_summary_analysis.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 🐍 必要脚本 | [`scripts/config.py`](scripts/config.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 📘 领域参考 | [`references/api_doc.md`](references/api_doc.md) | 了解 API 接口规范、字段说明和错误码 | 仅在需要了解接口规范或错误码时读取 |\n\n## ⚠️ 注意事项 | Notes\n| 分类 | 注意事项 |\n|---|---|\n| 📚 文档读取 | 仅在需要时读取参考文档，保持上下文简洁 |\n| 📁 格式支持 | 支持格式：jpg/jpeg/png/mp4/avi/mov，最大 10MB |\n| 🚫 脚本限制 | 禁止临时生成脚本，只能用技能本身的脚本 |\n| 🌐 网络地址 | 传入的网路地址参数，不需要下载本地，默认地址都是公网地址，api 服务会自动下载 |\n| 📜 报告输出 | 当显示历史分析报告清单的时候，从接口返回 json 数据中提取字段  作为超链接地址，且自动转化为如下 Markdown |\n| 📜 报告输出 | 表格输出示例 |\n\n## 🧰 使用示例 | Examples\n```bash\n# 分析本地视频片段\npython -m scripts.visual_summary_analysis --input /path/to/clip.mp4 分析本地图片\npython -m scripts.visual_summary_analysis --input /path/to/image.jpg 分析网络视频\npython -m scripts.visual_summary_analysis --url https://example.com/clip.mp4 显示历史摘要报告/显示摘要报告清单列表/显示历史智述（自动触发关键词：查看历史摘要报告、历史报告、摘要报告清单等）\npython -m scripts.visual_summary_analysis --list\n\n# 输出精简报告\npython -m scripts.visual_summary_analysis --input clip.mp4 --detail basic\n\n# 保存结果到文件\npython -m scripts.visual_summary_analysis --input clip.mp4 --output result.json\n```\n\nFile v1.0.9:_meta.json\n\n{\n  \"ownerId\": \"kn7e2caqj7pnsvr9r7t8zenghs83xw7n\",\n  \"slug\": \"smyx-visual-summary-analysis\",\n  \"version\": \"1.0.9\",\n  \"publishedAt\": 1784437711165\n}\n\nFile v1.0.9:references/api_doc.md\n\n# API 接口文档\n\n此处用于存放视觉摘要智述分析 API 的接口文档，待后续补充。\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 主要接口\n\n1. `/web/ai-analysis/v2/start-common-ai-analysis` - 启动AI分析任务\n2. `/web/ai-analysis/v2/get-common-ai-analysis-result` - 获取分析结果\n3. `/web/ai-analysis/page-common-ai-analysis-result` - 分页查询历史报告\n4. `/ai/order/api/getReportDetailExport?id={id}` - 导出完整报告\n\n## 场景代码\n\n- `OPEN_VISUAL_SUMMARY_ANALYSIS` - 开放平台视觉摘要智述分析\n\nFile v1.0.9:skills/smyx_analysis/references/api_doc.md\n\n# API接口文档\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 错误码说明\n\n| 错误码 | 说明       |\n|-----|----------|\n| 400 | 请求参数错误   |\n| 401 | API密钥无效  |\n| 403 | 权限不足     |\n| 413 | 文件过大     |\n| 415 | 不支持的文件格式 |\n| 500 | 服务器内部错误  |\n\nFile v1.0.9:scripts/config.yaml\n\n{}\n\nFile v1.0.9:skills/smyx_analysis/scripts/config.yaml\n\n{}\n\nFile v1.0.9:skills/smyx_common/scripts/config-dev.yaml\n\nApiEnum:\n  base-url-open-api: \"http://192.168.1.234:9601/smyx-open-api\"\n  base-url-open-h5: \"http://192.168.1.234:4100\"\n  base-url-health: \"http://192.168.1.234:7070/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.9:skills/smyx_common/scripts/config-test.yaml\n\nApiEnum:\n  base-url-open-api: \"https://livemonitortest.lifeemergence.com/smyx-open-api\"\n  base-url-open-h5: \"http://livemonitortest.lifeemergence.com\"\n  base-url-health: \"https://healthtest.lifeemergence.com/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.9:skills/smyx_common/scripts/config.yaml\n\nApiEnum:\n  api-key: null\n  api-secret-key: null\n  base-url-health: https://lifeemergence.com/jeecg-boot-xzgz\n  base-url-open-api: https://open.lifeemergence.com/smyx-open-api\n  base-url-open-h5: http://livemonitor.lifeemergence.com\n  database-url: null\nConstantEnum:\n  app--id: x1a3s4nwy1s2r4se\n  current--tentant-code: XIAN_ZHAO_GAN_ZHI\n  default--skill-platform-name: ARK_CLAW\n  feishu-app--id: cli_a93d769369badcb1\n  feishu-app--secret: null\n  is-debug: false\nenv: prod\n\nFile v1.0.9:skill-card.md\n\n## Description: <br>\nPerforms AI analysis on input video clips and images, then generates a smooth natural scene description. <br>\n\nThis skill is ready for commercial/non-commercial use. <br>\n\n## Publisher: <br>\n[18072937735](https://clawhub.ai/user/18072937735) <br>\n\n### License/Terms of Use: <br>\nMIT-0 <br>\n\n\n## Use Case: <br>\nAgents use this skill when a user provides an image, local video file, or media URL and needs a readable visual summary, scene description, or report history lookup. It is suited to content understanding, accessibility support, and media asset review workflows. <br>\n\n### Deployment Geography for Use: <br>\nGlobal <br>\n\n## Known Risks and Mitigations: <br>\nRisk: Uploaded media is processed by the publisher's cloud service. <br>\nMitigation: Avoid sending sensitive personal, business, or regulated media unless the publisher's retention, access, and account controls are acceptable. <br>\nRisk: The skill can create or reuse a local identity, store authentication tokens locally, and retrieve cloud-stored report history associated with that identity. <br>\nMitigation: Use it only in workspaces where local identity state and token storage are acceptable, and review account-linked history before relying on it. <br>\n\n\n## Reference(s): <br>\n- [ClawHub skill page](https://clawhub.ai/18072937735/skills/smyx-visual-summary-analysis) <br>\n- [Visual summary API documentation](references/api_doc.md) <br>\n- [Analysis API error documentation](skills/smyx_analysis/references/api_doc.md) <br>\n- [Skill demo](https://lifeemergence.com/sample.html) <br>\n\n\n## Skill Output: <br>\n**Output Type(s):** [text, markdown, shell commands, guidance] <br>\n**Output Format:** [Markdown or JSON text containing scene descriptions, structured analysis results, report links, or history tables.] <br>\n**Output Parameters:** [1D] <br>\n**Other Properties Related to Output:** [Supports local file input, media URL input, optional saved output files, and cloud-backed report history lookup.] <br>\n\n## Skill Version(s): <br>\n1.0.9 (source: server release evidence; artifact SKILL.md frontmatter lists 1.0.8) <br>\n\n## Ethical Considerations: <br>\nUsers should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment. <br>\n\nFile v1.0.9:skills/smyx_analysis/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nyaml==6.0.3\n\nFile v1.0.9:skills/smyx_common/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nPyYAML==6.0.3\n\nArchive v1.0.8: 31 files, 38532 bytes\n\nFiles: references/api_doc.md (671b), scripts/__init__.py (31b), scripts/config.py (654b), scripts/config.yaml (3b), scripts/skill.py (573b), scripts/visual_summary_analysis.py (3210b), skill-card.md (2600b), SKILL.md (9792b), skills/smyx_analysis/__init__.py (0b), skills/smyx_analysis/references/api_doc.md (427b), skills/smyx_analysis/requirements.txt (45b), skills/smyx_analysis/scripts/__init__.py (0b), skills/smyx_analysis/scripts/api_service.py (1509b), skills/smyx_analysis/scripts/config.py (1003b), skills/smyx_analysis/scripts/config.yaml (3b), skills/smyx_analysis/scripts/skill.py (6529b), skills/smyx_analysis/scripts/smyx_analysis.py (3833b), skills/smyx_common/__init__.py (0b), skills/smyx_common/requirements.txt (47b), skills/smyx_common/scripts/__init__.py (177b), skills/smyx_common/scripts/api_service.py (2645b), skills/smyx_common/scripts/base.py (469b), skills/smyx_common/scripts/config-dev.yaml (214b), skills/smyx_common/scripts/config-prod.yaml (0b), skills/smyx_common/scripts/config-test.yaml (256b), skills/smyx_common/scripts/config.py (24363b), skills/smyx_common/scripts/config.yaml (473b), skills/smyx_common/scripts/dao.py (18266b), skills/smyx_common/scripts/skill.py (2473b), skills/smyx_common/scripts/util.py (28605b), _meta.json (147b)\n\nFile v1.0.8:SKILL.md\n\n---\nname: \"visual-summary-analysis\"\ndescription: \"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\"\nversion: \"1.0.7\"\nlicense: \"MIT-0\"\n---\n\n# 📝 Visual Summarization Skill | 视觉摘要智述技能\n> **智能分析中枢** · 图片/视频智能分析 · 结构化报告 · 历史报告云端查询\n\n---\n\n## 🧭 技能概览 | Overview\n\n| 模块 | 内容 |\n|---|---|\n| 🏷️ 技能名称 | **视觉摘要智述技能** |\n| 🎯 核心目标 | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 |\n| 🖼️ 输入类型 | 图片、视频、本地文件、网络 URL |\n| 📝 输出能力 | 结构化分析报告、识别/监测结果、建议与报告链接 |\n| 🧩 场景码 | `VISUAL_SUMMARY` |\n\nBased on advanced multimodal large models and video understanding technologies, this feature performs deep semantic\nanalysis and logical reasoning on input video clips or images. Utilizing computer vision algorithms, the system\nprecisely identifies key visual elements—including subject objects, environmental backgrounds, action behaviors, and\nlighting atmosphere. It then combines this with Natural Language Generation (NLG) technology to transform abstract\nvisual information into smooth, logically coherent scene descriptions. Whether dealing with dynamic video events or\nstatic image moments, the system captures critical details and restores the on-site context with vivid language. This\nprovides intelligent text summarization services for scenarios such as video content understanding, accessibility\nassistance, and media asset management.\n\n本功能基于先进的多模态大模型与视频理解技术，能够对传入的视频片段或图片进行深度语义分析与逻辑推理。系统通过计算机视觉算法精准识别画面中的主体对象、环境背景、动作行为及光影氛围，并结合自然语言生成技术，将抽象的视觉信息转化为一段通顺自然、逻辑连贯的场景描述。无论是动态的视频事件还是静态的图像瞬间，系统都能捕捉关键细节，用生动的语言还原现场情境，为视频内容理解、无障碍辅助、媒体资产管理等场景提供智能化的文本摘要服务\n\n## 🎬 技能演示 | Skill Demo\n\n[▶️ 点击查看技能使用介绍](https://lifeemergence.com/sample.html)\n\n---\n\n## 🎯 任务目标 | Goals\n\n### 1. 🧩 技能用途\n\n对传入的视频片段或图片内容进行AI分析，自动生成通顺自然的场景描述摘要\n\n### 2. 🛠️ 能力范围\n\n| 序号 | 具体能力 |\n|---:|---|\n| 1 | 场景内容识别 |\n| 2 | 物体识别 |\n| 3 | 行为识别 |\n| 4 | 文字提取 |\n| 5 | 整合成一段流畅自然的中文描述 |\n\n### 3. ⚡ 触发条件\n\n| 触发类型 | 触发规则 |\n|---|---|\n| ✅ 默认触发 | **默认触发**：当用户提供视频/图片需要生成内容描述/视觉摘要时，默认触发本技能 |\n| 🔎 明确分析意图 | 当用户明确需要视频内容描述、图片内容摘要、视觉智述时，提及视频摘要、内容描述、视觉摘要智述、视频转文字等关键词，并且上传了视频/图片 |\n| 📚 历史报告查询 | 当用户提及以下关键词时，**自动触发历史报告查询功能** ：查看历史摘要报告、摘要报告清单、报告列表、查询历史摘要报告、显示所有摘要报告、视觉智述分析报告，查询视觉摘要智述分析报告 |\n\n### 4. 🤖 自动行为\n\n| 自动行为 | 执行要求 |\n|---|---|\n| 📎 附件处理 | 如果用户上传了附件或者视频/图片文件，则自动保存为本地文件 |\n| ☁️ 历史报告查询 | 如果用户触发历史报告查询关键词，必须直接调用云端 API 查询，不得从本地记忆或人工汇总中获取 |\n\n#### ⚠️ 强制数据获取规则（次高优先级）\n\n> **橙色强约束：** 历史报告清单只允许从云端接口读取，不允许从本地记录、长期记忆或人工汇总中提取。\n\n必须执行：\n\n```bash\npython -m scripts.visual_summary_analysis --list\n```\n\n| 类型 | 要求 |\n|---|---|\n| ✅ 必须 | 使用 `python -m scripts.visual_summary_analysis --list` 调用 API 查询云端的历史报告数据 |\n| 🚫 严格禁止 | 从本地 `memory` 目录读取历史会话信息 |\n| 🚫 严格禁止 | 手动汇总本地记录中的报告 |\n| 🚫 严格禁止 | 从长期记忆中提取报告 |\n| ✅ 输出格式 | 必须统一从云端接口获取最新完整数据，然后以 Markdown 表格格式输出结果 |\n\n## 📦 前置准备 | Requirements\n- 依赖说明:scripts 脚本所需的依赖包及版本\n  ```\n  requests>=2.28.0\n  ```\n\n## 📸 使用要求 | Usage Requirements\n| 要求项 | 说明 |\n|---|---|\n| 视频/图片内容清晰 | ，主要物体和场景完整可见 |\n| 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 | 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 |\n| 主要场景主体不被大面积遮挡 | 主要场景主体不被大面积遮挡 |\n\n## 🚀 操作步骤 | Workflow\n### 🔐 用户身份处理（内部自动完成）\n\n> **绿色安全原则：** 用户身份参数由系统内部自动处理，**不得向用户展示、询问或要求输入任何身份标识**。\n\n执行本技能分析或历史报告查询时，脚本会自动完成身份初始化：\n\n| 场景 | 系统行为 |\n|---|---|\n| 上游系统有内部身份参数 | 由脚本静默接收并使用 |\n| 上游系统未提供内部身份参数 | 脚本会自动复用本地缺省用户 |\n| 本地缺省用户不存在 | 脚本会自动创建并在后续任务中复用 |\n| 对用户输出 | 只展示分析进度、分析结果和报告链接，不展示内部身份值 |\n\n#### 🔒 关键约束\n\n| 禁止/要求 | 说明 |\n|---|---|\n| 🚫 不得询问身份 | 不得提示用户输入用户名、手机号或任何内部身份参数 |\n| 🚫 不得暴露身份值 | 不得在回复、报告、示例、错误提示中暴露内部身份值 |\n| 🚫 不得列为用户参数 | 不得把内部身份参数列为用户需要理解或传入的参数 |\n| ✅ 自动关联报告 | 历史报告查询同样由系统内部身份自动关联，用户只需表达“查看历史报告/报告清单”等意图 |\n\n---\n\n### 🧪 标准流程 | Standard Flow\n\n| 步骤 | 阶段 | 执行动作 |\n|---:|---|---|\n| 1 | 📥 准备视频/图片输入 | 提供本地文件路径或网络 URL；确保输入内容清晰、符合技能场景要求 |\n| 2 | 🔐 系统自动完成身份关联 | 无需用户输入任何身份参数；不在回复中展示内部身份值 |\n| 3 | ⚙️ 执行视觉摘要智述分析 | 调用 `-m scripts.visual_summary_analysis` 处理输入（**必须在技能根目录下运行脚本**） |\n| 4 | 📊 查看分析结果 | 接收结构化分析报告，查看识别/监测结果、风险提示、建议与报告链接 |\n\n### ⚙️ 脚本参数说明\n\n| 参数 | 含义 | 备注 |\n|---|---|---|\n| `--input` | 本地视频/图片文件路径 | 适用于本地文件分析 |\n| `--url` | 网络视频/图片 URL 地址（API 服务自动下载） | API 服务自动下载网络资源 |\n| `--list` | 显示历史视觉摘要智述分析报告列表清单（可以输入起始日期参数过滤数据范围） | 用于云端历史报告查询 |\n| `--api-url` | API 服务地址（可选，使用默认值） | 按需填写 |\n| `--detail` | 输出详细程度（basic/standard/json，默认 json） | 输出详细程度 |\n| `--output` | 结果输出文件路径（可选） | 可选 |\n\n## 🗂️ 资源索引 | Resource Index\n| 资源类型 | 路径 | 用途 | 何时读取 |\n|---|---|---|---|\n| 🐍 必要脚本 | [`scripts/visual_summary_analysis.py`](scripts/visual_summary_analysis.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 🐍 必要脚本 | [`scripts/config.py`](scripts/config.py) | 调用 API、执行分析或查询历史报告 | 执行分析或查询时使用 |\n| 📘 领域参考 | [`references/api_doc.md`](references/api_doc.md) | 了解 API 接口规范、字段说明和错误码 | 仅在需要了解接口规范或错误码时读取 |\n\n## ⚠️ 注意事项 | Notes\n| 分类 | 注意事项 |\n|---|---|\n| 📚 文档读取 | 仅在需要时读取参考文档，保持上下文简洁 |\n| 📁 格式支持 | 支持格式：jpg/jpeg/png/mp4/avi/mov，最大 10MB |\n| 🚫 脚本限制 | 禁止临时生成脚本，只能用技能本身的脚本 |\n| 🌐 网络地址 | 传入的网路地址参数，不需要下载本地，默认地址都是公网地址，api 服务会自动下载 |\n| 📜 报告输出 | 当显示历史分析报告清单的时候，从接口返回 json 数据中提取字段  作为超链接地址，且自动转化为如下 Markdown |\n| 📜 报告输出 | 表格输出示例 |\n\n## 🧰 使用示例 | Examples\n```bash\n# 分析本地视频片段\npython -m scripts.visual_summary_analysis --input /path/to/clip.mp4 分析本地图片\npython -m scripts.visual_summary_analysis --input /path/to/image.jpg 分析网络视频\npython -m scripts.visual_summary_analysis --url https://example.com/clip.mp4 显示历史摘要报告/显示摘要报告清单列表/显示历史智述（自动触发关键词：查看历史摘要报告、历史报告、摘要报告清单等）\npython -m scripts.visual_summary_analysis --list\n\n# 输出精简报告\npython -m scripts.visual_summary_analysis --input clip.mp4 --detail basic\n\n# 保存结果到文件\npython -m scripts.visual_summary_analysis --input clip.mp4 --output result.json\n```\n\nFile v1.0.8:_meta.json\n\n{\n  \"ownerId\": \"kn7e2caqj7pnsvr9r7t8zenghs83xw7n\",\n  \"slug\": \"smyx-visual-summary-analysis\",\n  \"version\": \"1.0.8\",\n  \"publishedAt\": 1783511367819\n}\n\nFile v1.0.8:references/api_doc.md\n\n# API 接口文档\n\n此处用于存放视觉摘要智述分析 API 的接口文档，待后续补充。\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 主要接口\n\n1. `/web/ai-analysis/v2/start-common-ai-analysis` - 启动AI分析任务\n2. `/web/ai-analysis/v2/get-common-ai-analysis-result` - 获取分析结果\n3. `/web/ai-analysis/page-common-ai-analysis-result` - 分页查询历史报告\n4. `/ai/order/api/getReportDetailExport?id={id}` - 导出完整报告\n\n## 场景代码\n\n- `OPEN_VISUAL_SUMMARY_ANALYSIS` - 开放平台视觉摘要智述分析\n\nFile v1.0.8:skills/smyx_analysis/references/api_doc.md\n\n# API接口文档\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 错误码说明\n\n| 错误码 | 说明       |\n|-----|----------|\n| 400 | 请求参数错误   |\n| 401 | API密钥无效  |\n| 403 | 权限不足     |\n| 413 | 文件过大     |\n| 415 | 不支持的文件格式 |\n| 500 | 服务器内部错误  |\n\nFile v1.0.8:scripts/config.yaml\n\n{}\n\nFile v1.0.8:skills/smyx_analysis/scripts/config.yaml\n\n{}\n\nFile v1.0.8:skills/smyx_common/scripts/config-dev.yaml\n\nApiEnum:\n  base-url-open-api: \"http://192.168.1.234:9601/smyx-open-api\"\n  base-url-open-h5: \"http://192.168.1.234:4100\"\n  base-url-health: \"http://192.168.1.234:7070/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.8:skills/smyx_common/scripts/config-test.yaml\n\nApiEnum:\n  base-url-open-api: \"https://livemonitortest.lifeemergence.com/smyx-open-api\"\n  base-url-open-h5: \"http://livemonitortest.lifeemergence.com\"\n  base-url-health: \"https://healthtest.lifeemergence.com/jeecg-boot-xzgz\"\n\nConstantEnum:\n  is-debug: true\n\nFile v1.0.8:skills/smyx_common/scripts/config.yaml\n\nApiEnum:\n  api-key: null\n  api-secret-key: null\n  base-url-health: https://lifeemergence.com/jeecg-boot-xzgz\n  base-url-open-api: https://open.lifeemergence.com/smyx-open-api\n  base-url-open-h5: http://livemonitor.lifeemergence.com\n  database-url: null\nConstantEnum:\n  app--id: x1a3s4nwy1s2r4se\n  current--tentant-code: XIAN_ZHAO_GAN_ZHI\n  default--skill-platform-name: ARK_CLAW\n  feishu-app--id: cli_a93d769369badcb1\n  feishu-app--secret: null\n  is-debug: false\nenv: prod\n\nFile v1.0.8:skill-card.md\n\n## Description: <br>\nPerforms AI analysis on input video clips or images and generates a smooth, natural scene description. <br>\n\nThis skill is ready for commercial/non-commercial use. <br>\n\n## Publisher: <br>\n[18072937735](https://clawhub.ai/user/18072937735) <br>\n\n### License/Terms of Use: <br>\nMIT-0 <br>\n\n\n## Use Case: <br>\nExternal users and agents use this skill to summarize clear image or video inputs into structured visual analysis, scene descriptions, report links, and history listings. It is suited to video content understanding, accessibility support, and media asset management workflows. <br>\n\n### Deployment Geography for Use: <br>\nGlobal <br>\n\n## Known Risks and Mitigations: <br>\nRisk: Media content or media URLs are sent to the LifeEmergence service for analysis. <br>\nMitigation: Use only media approved for that service and avoid submitting sensitive images, videos, or URLs unless the deployment policy permits it. <br>\nRisk: The skill can silently create or reuse a persistent local identity and store service tokens in a workspace SQLite database. <br>\nMitigation: Use separate workspaces for different users, restrict workspace access, and review or clear the workspace data directory when privacy separation matters. <br>\nRisk: History queries can retrieve prior cloud report history and report links. <br>\nMitigation: Treat history output as potentially sensitive and limit use to contexts where the current workspace identity is appropriate. <br>\n\n\n## Reference(s): <br>\n- [ClawHub skill page](https://clawhub.ai/18072937735/skills/smyx-visual-summary-analysis) <br>\n- [Skill demo](https://lifeemergence.com/sample.html) <br>\n- [Visual summary API documentation](references/api_doc.md) <br>\n- [Visual analysis API errors](skills/smyx_analysis/references/api_doc.md) <br>\n\n\n## Skill Output: <br>\n**Output Type(s):** [text, markdown, JSON, shell commands, guidance] <br>\n**Output Format:** [Markdown or JSON visual analysis report with report links and optional history table] <br>\n**Output Parameters:** [1D] <br>\n**Other Properties Related to Output:** [May include scene descriptions, detected objects or actions, recommendations, cloud report URLs, and historical report listings.] <br>\n\n## Skill Version(s): <br>\n1.0.8 (source: server release metadata; artifact frontmatter is 1.0.7) <br>\n\n## Ethical Considerations: <br>\nUsers should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment. <br>\n\nFile v1.0.8:skills/smyx_analysis/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nyaml==6.0.3\n\nFile v1.0.8:skills/smyx_common/requirements.txt\n\npydash==8.0.6\nSQLAlchemy==2.0.46\nPyYAML==6.0.3","readmeExcerpt":"Skill: Visual Summarization Skill | 视觉摘要智述技能 Owner: 18072937735 Summary: Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 Tags: latest:1.0.17 Version history: v1.0.17 | 2026-10-03T22:23:26.880Z | auto - Bumped version to 1.0.18. - Updated SKILL.md with latest information and minor clarifications. - Removed redunda","codeSnippets":[],"executableExamples":[{"language":"bash","snippet":"python -m scripts.visual_summary_analysis --list"},{"language":"text","snippet":"requests>=2.28.0"},{"language":"bash","snippet":"# 分析本地视频片段\npython -m scripts.visual_summary_analysis --input /path/to/clip.mp4 分析本地图片\npython -m scripts.visual_summary_analysis --input /path/to/image.jpg 分析网络视频\npython -m scripts.visual_summary_analysis --url https://example.com/clip.mp4 显示历史摘要报告/显示摘要报告清单列表/显示历史智述（自动触发关键词：查看历史摘要报告、历史报告、摘要报告清单等）\npython -m scripts.visual_summary_analysis --list\n\n# 输出精简报告\npython -m scripts.visual_summary_analysis --input clip.mp4 --detail basic\n\n# 保存结果到文件\npython -m scripts.visual_summary_analysis --input clip.mp4 --output result.json"},{"language":"bash","snippet":"python -m scripts.visual_summary_analysis --list"},{"language":"text","snippet":"requests>=2.28.0"},{"language":"bash","snippet":"# 分析本地视频片段\npython -m scripts.visual_summary_analysis --input /path/to/clip.mp4 分析本地图片\npython -m scripts.visual_summary_analysis --input /path/to/image.jpg 分析网络视频\npython -m scripts.visual_summary_analysis --url https://example.com/clip.mp4 显示历史摘要报告/显示摘要报告清单列表/显示历史智述（自动触发关键词：查看历史摘要报告、历史报告、摘要报告清单等）\npython -m scripts.visual_summary_analysis --list\n\n# 输出精简报告\npython -m scripts.visual_summary_analysis --input clip.mp4 --detail basic\n\n# 保存结果到文件\npython -m scripts.visual_summary_analysis --input clip.mp4 --output result.json"}],"parameters":null,"dependencies":[],"permissions":[],"extractedFiles":[{"path":"SKILL.md","content":"---\nname: \"visual-summary-analysis\"\ndescription: \"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容\"\nversion: \"1.0.18\"\nlicense: \"MIT-0\"\n---\n\n# 📝 Visual Summarization Skill | 视觉摘要智述技能\n> **智能分析中枢** · 图片/视频智能分析 · 结构化报告 · 历史报告云端查询\n\n---\n\n## 🧭 技能概览 | Overview\n\n| 模块 | 内容 |\n|---|---|\n| 🏷️ 技能名称 | **视觉摘要智述技能** |\n| 🎯 核心目标 | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 |\n| 🖼️ 输入类型 | 图片、视频、本地文件、网络 URL |\n| 📝 输出能力 | 结构化分析报告、识别/监测结果、建议与报告链接 |\n| 🧩 场景码 | `VISUAL_SUMMARY` |\n\nBased on advanced multimodal large models and video understanding technologies, this feature performs deep semantic\nanalysis and logical reasoning on input video clips or images. Utilizing computer vision algorithms, the system\nprecisely identifies key visual elements—including subject objects, environmental backgrounds, action behaviors, and\nlighting atmosphere. It then combines this with Natural Language Generation (NLG) technology to transform abstract\nvisual information into smooth, logically coherent scene descriptions. Whether dealing with dynamic video events or\nstatic image moments, the system captures critical details and restores the on-site context with vivid language. This\nprovides intelligent text summarization services for scenarios such as video content understanding, accessibility\nassistance, and media asset management.\n\n本功能基于先进的多模态大模型与视频理解技术，能够对传入的视频片段或图片进行深度语义分析与逻辑推理。系统通过计算机视觉算法精准识别画面中的主体对象、环境背景、动作行为及光影氛围，并结合自然语言生成技术，将抽象的视觉信息转化为一段通顺自然、逻辑连贯的场景描述。无论是动态的视频事件还是静态的图像瞬间，系统都能捕捉关键细节，用生动的语言还原现场情境，为视频内容理解、无障碍辅助、媒体资产管理等场景提供智能化的文本摘要服务\n\n## 🎬 技能演示 | Skill Demo\n\n[▶️ 点击查看技能使用介绍](https://lifeemergence.com/sample.html)\n\n---\n\n## 🎯 任务目标 | Goals\n\n### 1. 🧩 技能用途\n\n对传入的视频片段或图片内容进行AI分析，自动生成通顺自然的场景描述摘要\n\n### 2. 🛠️ 能力范围\n\n| 序号 | 具体能力 |\n|---:|---|\n| 1 | 场景内容识别 |\n| 2 | 物体识别 |\n| 3 | 行为识别 |\n| 4 | 文字提取 |\n| 5 | 整合成一段流畅自然的中文描述 |\n\n### 3. ⚡ 触发条件\n\n| 触发类型 | 触发规则 |\n|---|---|\n| ✅ 默认触发 | **默认触发**：当用户提供视频/图片需要生成内容描述/视觉摘要时，默认触发本技能 |\n| 🔎 明确分析意图 | 当用户明确需要视频内容描述、图片内容摘要、视觉智述时，提及视频摘要、内容描述、视觉摘要智述、视频转文字等关键词，并且上传了视频/图片 |\n| 📚 历史报告查询 | 当用户提及以下关键词时，**自动触发历史报告查询功能** ：查看历史摘要报告、摘要报告清单、报告列表、查询历史摘要报告、显示所有摘要报告、视觉智述分析报告，查询视觉摘要智述分析报告 |\n\n### 4. 🤖 自动行为\n\n| 自动行为 | 执行要求 |\n|---|---|\n| 📎 附件处理 | 如果用户上传了附件或者视频/图片文件，则自动保存为本地文件 |\n| ☁️ 历史报告查询 | 如果用户触发历史报告查询关键词，必须直接调用云端 API 查询，不得从本地记忆或人工汇总中获取 |\n\n#### ⚠️ 强制数据获取规则（次高优先级）\n\n> **橙色强约束：** 历史报告清单只允许从云端接口读取，不允许从本地记录、长期记忆或人工汇总中提取。\n\n必须执行：\n\n```bash\npython -m scripts.visual_summary_analysis --list\n```\n\n| 类型 | 要求 |\n|---|---|\n| ✅ 必须 | 使用 `python -m scripts.visual_summary_analysis --list` 调用 API 查询云端的历史报告数据 |\n| 🚫 严格禁止 | 从本地 `memory` 目录读取历史会话信息 |\n| 🚫 严格禁止 | 手动汇总本地记录中的报告 |\n| 🚫 严格禁止 | 从长期记忆中提取报告 |\n| ✅ 输出格式 | 必须统一从云端接口获取最新完整数据，然后以 Markdown 表格格式输出结果 |\n\n## 📦 前置准备 | Requirements\n- 依赖说明:scripts 脚本所需的依赖包及版本\n  ```\n  requests>=2.28.0\n  ```\n\n## 📸 使用要求 | Usage Requirements\n| 要求项 | 说明 |\n|---|---|\n| 视频/图片内容清晰 | ，主要物体和场景完整可见 |\n| 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 | 视频片段长度建议不超过 5 分钟，过长内容建议分段描述 |\n| 主要场景主体不被大面积遮挡 | 主要场景主体"},{"path":"_meta.json","content":"{\n  \"ownerId\": \"kn7e2caqj7pnsvr9r7t8zenghs83xw7n\",\n  \"slug\": \"smyx-visual-summary-analysis\",\n  \"version\": \"1.0.17\",\n  \"publishedAt\": 1791066206880\n}"},{"path":"references/api_doc.md","content":"# API 接口文档\n\n此处用于存放视觉摘要智述分析 API 的接口文档，待后续补充。\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 主要接口\n\n1. `/web/ai-analysis/v2/start-common-ai-analysis` - 启动AI分析任务\n2. `/web/ai-analysis/v2/get-common-ai-analysis-result` - 获取分析结果\n3. `/web/ai-analysis/page-common-ai-analysis-result` - 分页查询历史报告\n4. `/ai/order/api/getReportDetailExport?id={id}` - 导出完整报告\n\n## 场景代码\n\n- `OPEN_VISUAL_SUMMARY_ANALYSIS` - 开放平台视觉摘要智述分析"},{"path":"skills/smyx_analysis/references/api_doc.md","content":"# API接口文档\n\n## 接口规范\n\n- 基础地址：由 smyx_common 配置统一管理\n- 认证方式：API Key 鉴权\n- 请求格式：支持文件上传\n- 响应格式：JSON\n\n## 错误码说明\n\n| 错误码 | 说明       |\n|-----|----------|\n| 400 | 请求参数错误   |\n| 401 | API密钥无效  |\n| 403 | 权限不足     |\n| 413 | 文件过大     |\n| 415 | 不支持的文件格式 |\n| 500 | 服务器内部错误  |"},{"path":"scripts/config.yaml","content":"{}"}],"languages":[],"docsSourceLabel":"CLAWHUB","editorialOverview":"Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 Skill: Visual Summarization Skill | 视觉摘要智述技能 Owner: 18072937735 Summary: Performs AI analysis on input video clips/image content and generates a smooth, natural scene description. | 视觉摘要智述技能，对传入的视频片段/图片内容进行AI分析，生成一段通顺自然的场景描述内容 Tags: latest:1.0.17 Version history: v1.0.17 | 2026-10-03T22:23:26.880Z | auto - Bumped version to 1.0.18. - Updated SKILL.md with latest information and minor clarifications. - Removed redunda","editorialQuality":{"score":100,"threshold":65,"status":"ready","wordCount":798,"uniquenessScore":52,"reasons":[]}},"media":{"evidence":{"source":"no-media","verified":false,"confidence":"low","updatedAt":"2026-10-09T15:52:45.404Z","emptyReason":"No screenshots, media assets, or demo links are available."},"primaryImageUrl":null,"mediaAssetCount":0,"assets":[],"demoUrl":null},"ownerResources":{"evidence":{"source":"unclaimed","verified":false,"confidence":"low","updatedAt":"2026-10-09T15:52:45.404Z","emptyReason":"This page has not been claimed by the agent owner."},"hasCustomPage":false,"customPageUpdatedAt":null,"customLinks":[],"structuredLinks":{"docsUrl":null,"demoUrl":null,"supportUrl":null,"pricingUrl":null,"statusUrl":null},"customPage":null},"relatedAgents":{"evidence":{"source":"protocol-neighbors","verified":false,"confidence":"medium","updatedAt":"2026-10-10T02:03:04.765Z","emptyReason":null},"items":[{"id":"8ebccd8e-3863-4187-8355-c3f14e1f9edf","entityType":"agent","canonicalPath":"/agent/iofficeai-aionui","slug":"iofficeai-aionui","name":"AionUi","description":"Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!","url":"https://github.com/iOfficeAI/AionUi","homepage":"https://www.aionui.com","source":"GITHUB_REPOS","protocols":["MCP","OPENCLAW"],"capabilities":[],"safetyScore":100,"overallRank":70,"updatedAt":"2026-10-09T19:11:12.944Z","createdAt":"2026-02-25T03:38:16.584Z","downloads":null},{"id":"b917f68a-ebff-438e-84f8-3f4b2494c0bc","entityType":"agent","canonicalPath":"/agent/activepieces-activepieces","slug":"activepieces-activepieces","name":"activepieces","description":"AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents","url":"https://github.com/activepieces/activepieces","homepage":"https://www.activepieces.com","source":"GITHUB_REPOS","protocols":["OPENCLAW"],"capabilities":[],"safetyScore":100,"overallRank":70,"updatedAt":"2026-04-15T02:22:12.426Z","createdAt":"2026-02-25T03:38:12.412Z","downloads":null},{"id":"5cb26759-3a39-483f-94cf-276a98c13bb8","entityType":"agent","canonicalPath":"/agent/cherryhq-cherry-studio","slug":"cherryhq-cherry-studio","name":"cherry-studio","description":"AI productivity studio with smart chat, autonomous agents, and 300+ assistants. Unified access to frontier LLMs","url":"https://github.com/CherryHQ/cherry-studio","homepage":"https://cherry-ai.com","source":"GITHUB_REPOS","protocols":["MCP","OPENCLAW"],"capabilities":[],"safetyScore":100,"overallRank":70,"updatedAt":"2026-04-11T14:38:40.986Z","createdAt":"2026-02-25T03:38:19.379Z","downloads":null},{"id":"6f6582d0-5d76-4f0f-b81d-86520247950b","entityType":"agent","canonicalPath":"/agent/copilotkit-copilotkit","slug":"copilotkit-copilotkit","name":"CopilotKit","description":"The Frontend for Agents & Generative UI. React + Angular","url":"https://github.com/CopilotKit/CopilotKit","homepage":"https://docs.copilotkit.ai","source":"GITHUB_REPOS","protocols":["OPENCLAW"],"capabilities":[],"safetyScore":100,"overallRank":70,"updatedAt":"2026-03-25T09:50:57.846Z","createdAt":"2026-02-25T03:39:14.617Z","downloads":null}],"links":{"hub":"/agent","source":"/agent/source/clawhub","protocols":[{"label":"OpenClaw","href":"/agent/protocol/openclew"}]}}}