Sgrep
Skill XiaoConstantine/sgrep
semantic grep
npx -y skills add XiaoConstantine/sgrepAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
One thing to look at
- 15 stars15 stars. Stars are a popularity signal and not a quality one, but at this level it is likely that nobody has read this closely except its author, and you would be relying on your own review.
What its author says it does
Copied from the file, not written here
Semantic and hybrid code and conversation search for intent-based queries. Use when exploring unfamiliar codebases, finding code by concept instead of exact text, or recalling past agent conversations about similar problems.
The file declares its own license as Apache-2.0. That is the author’s claim about this one file, and it is not the same thing as the license GitHub reports for the repository, which is listed with the other numbers below.
SKILL.md
4.0 KB, as published. Nobody here has run it
sgrep - Smart Code & Conversation Search
Use sgrep for semantic and hybrid search across code and agent conversations. It understands intent, not just exact strings.
When to Use
Code Search
- Finding code by concept: "error handling", "authentication logic", "rate limiting"
- Searching for specific terms with semantic context: use
--hybrid - Best code-search accuracy after indexing: use
--hybrid --colbert - Exploring unfamiliar codebases
- When ripgrep patterns keep missing relevant code
Conversation Search
- Finding past discussions with Claude Code, Codex CLI, Cursor, OpenCode, or Pi
- Recalling how you solved a similar problem before
- Building context from previous sessions for new tasks
- Searching across all your coding agent interactions
Commands
# First time only
sgrep setup
sgrep setup --with-rerank # optional, only for --rerank
# Index current directory; builds compact TQ-MSE chunk/file vectors by default
sgrep index .
# Optional ColBERT segment codec override
sgrep index . --colbert-codec tqmse
sgrep index . --colbert-codec int8
sgrep index . --colbert-codec pq6
# Legacy compatibility: also persist full SQL vectors
sgrep index . --sql-vectors
# Watch mode keeps SQL vectors for incremental updates; rerun index to compact
sgrep watch .
# Balanced semantic + lexical code search (default)
sgrep "database connection pooling"
sgrep "how are errors handled"
# Fast semantic-only search
sgrep --profile fast "error handling"
# Best code-search accuracy
sgrep --profile quality "JWT validation"
sgrep --profile quality "authentication middleware"
# With code context
sgrep -c "authentication middleware"
# JSON output
sgrep --json "rate limiting"
Conversation Search
# Index conversations; refreshes compact TQ-MSE turn vectors
sgrep conv index
sgrep conv index --source codex
sgrep conv index --source claude
sgrep conv index --source opencode
sgrep conv index --source pi
sgrep conv index --watch
sgrep conv index --force
# Search conversations
sgrep conv "authentication flow"
sgrep conv "JWT refresh_token" --hybrid
sgrep conv "database migration" --agent claude --since 7d
sgrep conv "bug fix" --project payment-service --after 2026-01-01 --before 2026-06-01
sgrep conv "exact phrase" --exact
sgrep conv "auth" --json -n 1
# View, export, context, and copy helpers
sgrep conv view <session_id>
sgrep conv view <session_id> --turn 3 --no-color
sgrep conv export <session_id> --format markdown -o conversation.md
sgrep conv export <session_id> --format json -o conversation.json
sgrep conv context <session_id>
sgrep conv context <session_id> --turns 10 --copy
sgrep conv copy <session_id> --turn 2 --code-only
sgrep conv status
Semantic vs Hybrid
| Mode | Best For | Example |
|---|---|---|
--profile fast | Lowest-latency semantic exploration | "how does auth work" |
--profile balanced (default) | Semantic + exact-term recall | "JWT token validation" |
--profile quality | Highest code-search accuracy | "authentication middleware" |
Use the default balanced profile for most agent searches and --profile quality when ranking quality matters more than minimum latency.
Search Hierarchy
- sgrep → Balanced semantic + lexical discovery
- sgrep --profile fast → Lowest-latency semantic discovery
- sgrep --profile quality → Rerank candidates with precomputed late interaction
- ast-grep → Match structural patterns in those files
- ripgrep → Exact text for specific symbols