agentsclimarketplace

Yao deepseek crawler

Skill ShenyuanNext/wepr-growth-skills/skills/yao-deepseek-crawler

Use when a user provides DeepSeek web AI-search keywords, repeat count, target entity, and entity type, then needs repeated fresh-window crawls aggregated into JSON plus a Kami HTML GEO report. Not for generic website crawling, DeepSeek API chat, SEO writing, or one-off answer generation.From its SKILL.md

Install
npx -y skills add ShenyuanNext/wepr-growth-skills --skill yao-deepseek-crawler

Assembled from the repository path, not quoted from the project. Check it against their README if it does not work.

One thing to look at

  • 2 stars2 stars. Stars are a popularity signal and not a quality one, but at this level it is likely that nobody has read this closely except its author, and you would be relying on your own review.

SKILL.md

2.4 KB, 468 tokens by cl100k_base, as published. Nobody here has run it

Yao DeepSeek Crawler

Inputs

Standard inputs: keywords/questions, repeat count, target entity, entity type (人/person, 公司/company, 产品/product), browser profile, and optional output directory. Competitors must match the target type. Reports default to Simplified Chinese with an English summary toggle.

Workflow

  1. Read references/user-setup-and-usage.md for install, prerequisites, and user-facing steps.
  2. Read references/deepseek-crawl-workflow.md for crawler setup, preflight, delay, resume, and batch rules.
  3. Read references/report-contract.md for JSON schema, metrics, target/competitor recognition, and report rules.
  4. Run node scripts/preflight.mjs --profile <profile> before fresh crawling.
  5. Stage 1: run scripts/deepseek_batch_crawl.mjs with questions, repeat, profile, target entity/type, --safe-random-delay, and output dir.
  6. Stage 2: run scripts/analyze_deepseek_results.py on any crawl JSON with target entity/type, optional brands file, report output dir, and semantic review mode. Use --semantic-review auto by default; use --semantic-review required for formal delivery when AI review must pass.
  7. Return the raw crawl JSON, structured Markdown, structured Excel workbook, HTML report, summary JSON, semantic-review cache when present, and failed logs. Reports include AI semantic labels for entity recognition, target-vs-best-3 radar, click-to-reveal bubbles, Chinese source names, clickable citations, title intent, compact treemap, and GEO actions.

Honest Boundaries

  • Do not use for generic website crawling, DeepSeek API chat, SEO copywriting, or one-off answer generation.
  • Reuses local DeepSeek web automation; does not bypass login, CAPTCHA, bot checks, or hidden data.
  • Probability metrics are repeated-sample estimates, not ground truth.
  • Inferred competitors are heuristic unless --semantic-review required passes. AI semantic review is an audit enhancement and never replaces hard-rule gates or answer-body evidence.
  • Review aliases, semantic labels, excluded candidates, and competitor tables before external use.
  • Preserve raw answers, reference titles, URLs, and logs.

Keep looking

Skills are one crate of 325,949. Ordering is by how many stacks a row turns up in, so the top of any crate is what has actually been picked rather than what has the most stars.