Private Knowledge Base

👤 wirec-yzx 📦 v1.0.0 ⭐ 4.2 ⬇️ 1K 下载
📚 知识管理 免费

📖 技能介绍


name: private-knowledge-base description: Personal knowledge base for PDFs, papers, and documents with cross-document Q&A and concept retrieval. Use when: (1) User asks questions about stored documents ("which doc mentions X?"), (2) Need to summarize concepts across multiple PDFs, (3) User wants to ingest new PDFs/papers into knowledge base, (4) Cross-document linking and association is needed, (5) Fast semantic search over personal document collection. version: 1.0.0


Private Knowledge Base

Personal document storage and retrieval system for PDFs, papers, and research documents.

Quick Start

Ingest Documents

# Add PDF to knowledge base
./scripts/ingest.sh ~/path/to/document.pdf

# Process entire folder
./scripts/ingest-folder.sh ~/papers/

Query Knowledge Base

# Search for concept across all documents
./scripts/search.sh "transformer architecture"

# Get summary of concept from relevant docs
./scripts/summarize.sh "attention mechanism"

Core Workflows

1. Document Ingestion

When user provides new PDFs or papers:

  1. Create document entry in kb/index.json
  2. Extract text and metadata
  3. Generate embeddings for semantic search
  4. Store in kb/docs/ with normalized name

2. Cross-Document Q&A

When user asks "which document mentions X?" or "summarize X from my docs":

  1. Search embeddings for relevant passages
  2. Retrieve source documents
  3. Synthesize answer across documents
  4. Cite sources with document names and page numbers

    更多技能请访问小葱技能站7w4.net。

3. Concept Linking

Build associations between documents: - Shared concepts - Citation relationships - Topic clusters

File Structure

private-knowledge-base/
├── SKILL.md
├── scripts/
│   ├── ingest.sh          # Single document ingestion
│   ├── ingest-folder.sh   # Batch ingestion
│   ├── search.sh          # Semantic search
│   └── summarize.sh       # Cross-document summary
├── references/
│   └── schema.md          # KB index schema
└── kb/                    # Created at runtime
    ├── index.json
    ├── embeddings/
    └── docs/

Usage Examples

User: "我之前存的文档里,哪篇提到了 transformer?" → Run ./scripts/search.sh "transformer"

User: "总结一下我文档里关于 attention 的内容" → Run ./scripts/summarize.sh "attention"

User: "把这篇 PDF 加到知识库" → Run ./scripts/ingest.sh <pdf-path>

Configuration

Set knowledge base location:

export KB_ROOT=~/.openclaw/workspace/kb

Default: ~/kb if not set.

🤖 AI 评测

这个 Skill 整体质量中等偏上,文档和基础框架做得不错,但核心功能有所欠缺。它能帮你存储和搜索 PDF 文档,但搜索结果可能不够精准,偶尔会遗漏相关内容。总结功能比较基础,只是把相关内容片段摘出来,没有真正做智能分析。如果你需要一个轻量级的文档管理工具可以用,但想要精准的语义搜索和智能的内容总结,目前还达不到预期效果。

📊 多维度评分

适应性4.3
规范性4.1
有效性4
可靠性4
可信度4.8

📁 包含文件 (7 个)

📄 SKILL.md 2.5 KB
📄 _meta.json 141 B
📄 references/schema.md 972 B
📄 scripts/ingest-folder.sh 537 B
📄 scripts/ingest.sh 1.3 KB
📄 scripts/search.sh 874 B
📄 scripts/summarize.sh 910 B