feat: keep per-item extracted outputs by default

This commit is contained in:
root
2026-03-28 23:14:18 +08:00
parent 0a00a97724
commit 9fd21c59f5
4 changed files with 18 additions and 6 deletions
+4 -1
View File
@@ -102,13 +102,16 @@ By default the pipeline writes only:
- `outputs/freshrss/rerun/<run_id>/raw/freshrss.raw.json`
- `outputs/freshrss/rerun/<run_id>/candidates/openclaw-delivery-payload.json`
- `outputs/freshrss/rerun/<run_id>/run-report.json`
- `outputs/freshrss/rerun/<run_id>/extracted/item-XX.extracted.json` (one per item)
It also updates local runtime keyword data:
- `data/term_index/daily/YYYY-MM-DD.json`
- `data/term_index/term_stats.json`
If `debug_artifacts=true`, the pipeline additionally writes per-item intermediate files.
Per-item extracted files live under `extracted/` and are always written.
If `debug_artifacts=true`, the pipeline additionally writes normalized items, summaries, filter decisions, candidate records, and candidate inputs.
The main pipeline does not emit a batch-level `freshrss.extracted.json` file by default.
## Payload Specs