Nature Literature Pipeline

作者 yuan1z0825d5c8baa15d6bMIT46K 个星标收录于 2026年10月8日更新于 2026年10月8日仓库今天更新

Complete automated literature discovery pipeline: multi-source search → six-dimension scoring → fine reading → formatted delivery → archival. Combines a configurable engine with daily cron-driven application layer. Works with Feishu, Telegram, or any messaging platform.

AI 生成的概览

自动化的每日文献发现流水线,完成检索、评分、精读、推送与归档。

功能
该技能定义了一套由定时任务驱动的结构化文献流水线:从 arXiv、OpenAlex、Crossref 和 Semantic Scholar 检索候选论文,用六维加权评分规则筛选,对排名靠前的文献进行精读,并生成摘要推送到飞书、Telegram 等消息平台。随后按 DOI 或 arXiv ID 去重、分类,并把标准化文献笔记写入归档目录。该技能仅包含说明文档,参考文件涵盖评分体系、研究空白分析、笔记模板、推送格式、定时任务配置和综述编写流程。
适用场景
适用于需要对特定研究领域进行长期、自动化文献监测而非一次性检索的场景。适合希望每天收到新论文排序摘要并持续积累去重笔记库的研究人员,也适合集中撰写文献综述的工作。
运行要求
该技能不附带脚本,只有说明与参考文档。需要智能体能在持续运行的机器上执行定时任务,需要访问文献 API(arXiv、OpenAlex、Crossref、Semantic Scholar)的网络权限,以及飞书群或 Telegram 频道等推送目标。关键词、评分权重、分类规则、推送目标和归档路径都需要配置。

Nature Literature Pipeline

A complete, production-tested automated literature pipeline. Not just "search for papers" — it's a structured engine that scores, classifies, reads, delivers, and archives research papers daily.

What It Does

Cron (daily trigger, e.g. 08:30)  │  ├─ ① SEARCH (30 candidates)  │   arXiv / OpenAlex / Crossref / Semantic Scholar (auto-degradation)  │  ├─ ② COARSE FILTER (30 → 5)  │   Six-dimension scoring: topic match × 35 + methodology × 20  │   + journal quality × 15 + network relevance × 10  │   + applied value × 10 + archival value × 10  │  ├─ ③ FINE READ (top 5)  │   Abstract-level or full-text. Source level tagged:  │   Full-text / Abstract only / Metadata only  │  ├─ ④ DELIVER  │   Formatted digest to Feishu/Telegram/etc.  │   🏅 rank | title | journal | ⭐ score | 💡 one-liner  │   🔬 methods | 📊 key results | 🧭 commentary  │  └─ ⑤ ARCHIVE      DOI/arXiv de-dup → classify → write notes → update index

Quick Start

After installing, tell your agent:

My research area is [X], keywords: [Y], deliver to [feishu group name], archive to [path]

The agent will configure keywords, delivery target, and archive path automatically.

Then set up a daily cron job:

Set up a daily literature push at 08:30 Beijing time, 30 candidates, top 5 delivered

Architecture

The skill is organized in two layers:

LayerPurposeFiles
EngineScoring, classification, note templates, gap analysisreferences/scoring-system.md, references/gap-analysis.md, references/note-template.md
ApplicationDaily cron pipeline, delivery formatting, archival workflowreferences/push-format.md, references/cron-setup.md, references/review-compilation-workflow.md

Configuration

All domain-specific content is configurable:

  • Keywords — your research keywords (English + Chinese)
  • Scoring weights — adjust the six dimensions for your field
  • Classification rules — define your own tier system (A-E or custom)
  • Delivery target — Feishu group, Telegram channel, email, etc.
  • Archive path — local vault/wiki directory

A config template is provided in templates/literature-push-template.md.

Built-in Safeguards

  • Score validation: Each dimension capped, total recalculated — no 11/10 allowed
  • Triple de-duplication: DOI / arXiv ID / OpenAlex ID
  • Graceful degradation: Semantic Scholar down → auto-switch to OpenAlex + Crossref + arXiv
  • Read-only archive: Daily pipeline writes to raw/ literature directory only; never modifies wiki/knowledge base without user approval

Related Skills

  • nature-academic-search — ad-hoc literature search (complementary; this skill adds structured daily automation)
  • nature-citation — CNS citation export (for importing pipeline discoveries into manuscripts)
  • zotero — library management (for long-term organization of pipeline outputs)
  • arxiv — arXiv API (used as a search source)

References

ReferencePurpose
references/scoring-system.mdSix-dimension scoring rubric with weights, caps, and evaluation logic
references/gap-analysis.mdMethodology for identifying research gaps through systematic literature survey
references/note-template.mdStandardized literature note format with YAML frontmatter
references/push-format.mdDaily digest message template with field guidelines and example
references/cron-setup.mdCron job creation, verification, and manual fallback procedures
references/review-compilation-workflow.mdEnd-to-end workflow for concentrated literature review writing

Pitfalls

  1. Keyword drift: Review keywords monthly — research directions evolve
  2. Score inflation: Subagents may inflate scores; always validate arithmetic
  3. Duplicate creep: Classic papers will reappear; maintain a dedup index
  4. Wiki safety: Pipeline writes to raw/ only; wiki integration is manual
  5. Cron locality: Hermes cron is local, not cloud — machine must be running

来源与署名

来源:yuan1z0825/nature-skills位于skills/nature-literature-pipeline提交d5c8baa

许可证: MIT

内容归原作者所有。SourceWeft 从公开仓库中收录这些内容。

举报或申请下架