Rag pipeline
全球最大的 Claude Code 技能聚合库 · 收录 3900+ 来自 12+ 来源的技能,提供在线搜索与趋势分析看板 / The world's largest Claude Code skill aggregation hub — 3900+ skills from 12+ sources with online search and trend dashboard
npx -y skills add bg-szy/TOP-SKILLS --skill rag-pipelineAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
2 things to look at
- no licenseNo license file was found in the repository. Code published without one is not open source by default, so using it at work is a question for whoever answers licensing questions where you are.
- 4 stars4 stars. Stars are a popularity signal and not a quality one, but at this level it is likely that nobody has read this closely except its author, and you would be relying on your own review.
What its author says it does
Copied from the file, not written here
Details on the Retrieval Augmented Generation pipeline, Ingestion, and Vector Search.
SKILL.md
1.0 KB, as published. Nobody here has run it
RAG Pipeline Logic
Ingestion
- Script:
backend/ingest.py - Process:
- Scans
docs/. - Cleans MDX (removes frontmatter/imports).
- Chunks text (1000 chars, 100 overlap).
- Embeds using
models/text-embedding-004. - Upserts to Qdrant collection
physical_ai_book.
- Scans
- Run:
python backend/ingest.py
Vector Search (Qdrant)
- Client:
qdrant-client - Collection:
physical_ai_book - Vector Size: 768 (Gecko-004)
- Similarity: Cosine
Prompt Engineering
- File:
backend/utils/helpers.py. - RAG Prompt: Constructs a prompt containing retrieved context chunks.
- Personalization:
backend/personalization.pycreates system instructions based onsoftware_backgroundandhardware_backgroundof the user.
Agentic Flow
We use a custom Agent class (backend/agents.py) that wraps the LLM calls, allowing for future expansion into multi-agent workflows.