Translate skill docs (article-deep-summary, llm-summary-review, keyword-cleanup-review) from English to Chinese

This commit is contained in:
wdm
2026-03-30 23:14:22 +08:00
parent 5b8df317ef
commit ca86705646
3 changed files with 158 additions and 158 deletions
+48 -48
View File
@@ -1,62 +1,62 @@
---
name: keyword-cleanup-review
description: Review and curate this repository's daily keyword index and frequency stats. Use when the user wants to inspect `data/term_index/term_stats.json`, recent `data/term_index/daily/*.json`, `configs/term_aliases.json`, `configs/term_stopwords.json`, or `configs/filter_context.personal.json` to propose alias merges, stopwords, watch terms, or `interest_keywords` updates without directly modifying configs.
description: 审查和整理本仓库的每日关键词索引和频率统计。当用户需要检查 `data/term_index/term_stats.json`、最近的 `data/term_index/daily/*.json`、`configs/term_aliases.json`、`configs/term_stopwords.json` 或 `configs/filter_context.personal.json`,以提议别名合并、停用词、关注词或 `interest_keywords` 更新(但不直接修改配置)时使用。
---
# Keyword Cleanup Review
# 关键词清理审查
Use this skill to turn the repository's keyword statistics into reviewable cleanup suggestions.
使用此技能将仓库的关键词统计转化为可审查的清理建议。
## Workflow
## 工作流程
1. Build a compact review bundle:
1. 构建精简的审查数据包:
```bash
python skills/keyword-cleanup-review/scripts/build_review_bundle.py
```
Optional knobs:
可选参数:
- `--days 7`
- `--top 50`
- `--output outputs/term_index/review/keyword-cleanup-bundle.json`
2. Read the generated bundle and the suggestion schema:
2. 阅读生成的数据包和建议模式:
- `outputs/term_index/review/keyword-cleanup-bundle.json`
- `skills/keyword-cleanup-review/references/suggestion-schema.md`
3. Produce two outputs:
3. 生成两份输出:
- A short Markdown review for humans
- A JSON suggestion file matching the schema
- 一份简短的供人工审阅的 Markdown 报告
- 一份符合模式的 JSON 建议文件
4. Keep the boundary strict:
4. 严格保持边界:
- Suggest changes to `configs/term_aliases.json`
- Suggest changes to `configs/term_stopwords.json`
- Suggest additions to `configs/filter_context.personal.json`
- Do not directly edit these files unless the user explicitly asks
- Do not suggest direct edits to `configs/filter_rules.json` unless the user asks for rule logic changes
- 建议 `configs/term_aliases.json` 的修改
- 建议 `configs/term_stopwords.json` 的修改
- 建议 `configs/filter_context.personal.json` 的新增
- 除非用户明确要求,否则不要直接编辑这些文件
- 除非用户要求修改规则逻辑,否则不要建议直接编辑 `configs/filter_rules.json`
## Review Heuristics
## 审查启发式规则
Prioritize these decisions:
优先考虑以下决策:
- Alias suggestion
- Same concept with different naming, casing, abbreviation, or Chinese/English variants
- Stopword suggestion
- Too generic, too broad, or too noisy to help filtering
- Interest keyword suggestion
- High-frequency and aligned with the user's backend engineering, AI-agent, and frontier-tech focus
- Watch term
- Recent and potentially important, but evidence is still weak
- 别名建议
- 同一概念的不同命名、大小写、缩写或中英文变体
- 停用词建议
- 过于通用、过于宽泛或噪声过大,对过滤无帮助
- 兴趣关键词建议
- 高频且与用户的后端工程、AI Agent 和前沿技术关注方向一致
- 关注词
- 近期出现且可能重要,但证据尚不充分
Prefer conservative suggestions. If confidence is low, put the term into `watch_terms`.
建议保守为主。如果置信度较低,将词放入 `watch_terms`。
## Inputs
## 输入
Primary inputs:
主要输入:
- `data/term_index/term_stats.json`
- `data/term_index/daily/*.json`
@@ -67,36 +67,36 @@ Primary inputs:
- `configs/term_watchlist.json`
- `configs/term_change_log.json`
The bundled script already compacts these into a single review bundle.
打包脚本已将这些内容压缩为单个审查数据包。
## Output Expectations
## 输出要求
The Markdown output should:
Markdown 输出应:
- Summarize the current state briefly
- List the top terms worth acting on
- Separate alias, stopword, interest-keyword, and watch-term recommendations
- Explain reasoning in short, concrete sentences
- 简要总结当前状态
- 列出值得处理的高频词
- 分类别名、停用词、兴趣关键词和关注词建议
- 用简短、具体的句子解释理由
The JSON output should follow:
JSON 输出应遵循:
- `references/suggestion-schema.md`
## Repository Notes
## 仓库说明
Current repository behavior:
当前仓库行为:
- Keyword stats are program-maintained, not LLM-maintained
- Stats are built from `keywords`, not `topics`
- Stats only include non-`drop` candidates
- `data/term_index/term_stats.json` is rebuilt from daily files, so reruns overwrite the same day instead of double-counting
- cleanup policy, watchlist, and change log are repository-managed governance inputs and should be respected during review
- 关键词统计由程序维护,而非 LLM 维护
- 统计基于 `keywords` 构建,而非 `topics`
- 统计仅包含非 `drop` 候选项
- `data/term_index/term_stats.json` 从每日文件重建,因此重新运行会覆盖同一天的数据而非重复计算
- 清理策略、关注列表和变更日志是仓库管理的治理输入,审查时应予以尊重
Keep suggestions aligned with that design.
保持建议与此设计保持一致。
## Resources
## 资源
- Script:
- 脚本:
- `scripts/build_review_bundle.py`
- Reference:
- 参考文档:
- `references/suggestion-schema.md`