Arxiv search
Search arXiv physics, math, and computer science preprints using natural language queries. Powered by Valyu semantic search.From its SKILL.md
npx -y skills add FridrichMethod/awesome-skills --skill arxiv-searchAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
2 things to look at
- no licenseNo license file was found in the repository. Code published without one is not open source by default, so using it at work is a question for whoever answers licensing questions where you are.
- 13 stars13 stars. Stars are a popularity signal and not a quality one, but at this level it is likely that nobody has read this closely except its author, and you would be relying on your own review.
What its file declares
Copied from the file, not written here
The file declares its own license as MIT. That is the author’s claim about this one file, and it is not the same thing as the license GitHub reports for the repository, which is listed with the other numbers below.
SKILL.md
5.4 KB, ~1.3k tokens by cl100k_base, as published. Nobody here has run it
arXiv Search
Search the complete arXiv database of preprints across physics, mathematics, computer science, and quantitative biology using natural language queries powered by Valyu's semantic search API.
Why This Skill is Powerful
- No API Parameter Parsing: Just pass natural language queries directly - no need to construct complex search parameters
- Semantic Search: Understands the meaning of your query, not just keyword matching
- Full-Text Access: Returns complete article content, not just abstracts
- Image Links: Includes figures and images from papers
- Comprehensive Coverage: Access to all of arXiv's preprint archive across multiple disciplines
Requirements
- Node.js 18+ (uses built-in fetch)
- Valyu API key from https://platform.valyu.ai ($10 free credits)
CRITICAL: Script Path Resolution
The scripts/search commands in this documentation are relative to this skill's installation directory.
Before running any command, locate the script using:
ARXIV_SCRIPT=$(find ~/.claude/plugins/cache -name "search" -path "*/arxiv-search/*/scripts/*" -type f 2>/dev/null | head -1)
Then use the full path for all commands:
$ARXIV_SCRIPT "quantum entanglement" 15
API Key Setup Flow
When you run a search and receive "setup_required": true, follow this flow:
-
Ask the user for their API key: "To search arXiv, I need your Valyu API key. Get one free ($10 credits) at https://platform.valyu.ai"
-
Once the user provides the key, run:
scripts/search setup <api-key> -
Retry the original search.
Example Flow:
User: Search arXiv for transformer architecture papers
→ Response: {"success": false, "setup_required": true, ...}
→ Claude asks: "Please provide your Valyu API key from https://platform.valyu.ai"
→ User: "val_abc123..."
→ Claude runs: scripts/search setup val_abc123...
→ Response: {"success": true, "type": "setup", ...}
→ Claude retries: scripts/search "transformer architecture papers" 10
→ Success!
When to Use This Skill
- Searching preprints across physics, mathematics, and computer science
- Finding research before peer review publication
- Cross-disciplinary research combining fields
- Staying current with rapid developments in AI and theoretical physics
- Prior art searching for new ideas
- Tracking emerging research trends
Output Format
{
"success": true,
"type": "arxiv_search",
"query": "quantum entanglement",
"result_count": 10,
"results": [
{
"title": "Article Title",
"url": "https://arxiv.org/abs/...",
"content": "Full article text with figures...",
"source": "arxiv",
"relevance_score": 0.95,
"images": ["https://example.com/figure1.jpg"]
}
],
"cost": 0.025
}
Processing Results
With jq
# Get article titles
scripts/search "query" 10 | jq -r '.results[].title'
# Get URLs
scripts/search "query" 10 | jq -r '.results[].url'
# Extract full content
scripts/search "query" 10 | jq -r '.results[].content'
Common Use Cases
AI/ML Research
# Find recent machine learning papers
scripts/search "large language model architectures" 50
Physics Research
# Search for quantum physics papers
scripts/search "topological quantum computation" 20
Mathematics
# Find math papers
scripts/search "representation theory and Lie algebras" 15
Computer Science
# Search for CS theory papers
scripts/search "distributed systems consensus algorithms" 25
Error Handling
All commands return JSON with success field:
{
"success": false,
"error": "Error message"
}
Exit codes:
0- Success1- Error (check JSON for details)
API Endpoint
- Base URL:
https://api.valyu.ai/v1 - Endpoint:
/search - Authentication: X-API-Key header
Architecture
scripts/
├── search # Bash wrapper
└── search.mjs # Node.js CLI
Direct API calls using Node.js built-in fetch(), zero external dependencies.
Adding to Your Project
If you're building an AI project and want to integrate arXiv Search directly into your application, use the Valyu SDK:
Python Integration
from valyu import Valyu
client = Valyu(api_key="your-api-key")
response = client.search(
query="your search query here",
included_sources=["valyu/valyu-arxiv"],
max_results=20
)
for result in response["results"]:
print(f"Title: {result['title']}")
print(f"URL: {result['url']}")
print(f"Content: {result['content'][:500]}...")
TypeScript Integration
import { Valyu } from "valyu-js";
const client = new Valyu("your-api-key");
const response = await client.search({
query: "your search query here",
includedSources: ["valyu/valyu-arxiv"],
maxResults: 20
});
response.results.forEach((result) => {
console.log(`Title: ${result.title}`);
console.log(`URL: ${result.url}`);
console.log(`Content: ${result.content.substring(0, 500)}...`);
});
See the Valyu docs for full integration examples and SDK reference.
What ships with it: 2 files
3.9 KB alongside SKILL.md, 2 of them executable
scripts/
- searchruns387 B
- search.mjsruns3.6 KB
Gives 0 of the 12 instructions most literature review skills give in ~1.3k tokens
Counted across 185 of the 191 authors here whose files we hold, read 2026-09-06
- Screen sources by title, abstract, then full textin 16 of 185, across 9 files
- Create a search protocol before collecting sourcesin 11 of 185, across 5 files
- Convert the prompt into a searchable research questionin 11 of 185, across 5 files
- Keep a reproducible search login 11 of 185, across 5 files
- Group evidence by theme, not per paperin 11 of 185, across 5 files
- Ask the user which review rigor level is neededin 11 of 185, across 5 files
- Use a structured extraction tablein 11 of 185, across 5 files
- Separate claims by confidence levelin 10 of 185, across 4 files
- Deduplicate by DOI, identifier, then titlein 10 of 185, across 4 files
- Record exclusion reasons at each screening stagein 9 of 185, across 3 files
- Verify citation identifiers before finalizingin 9 of 185, across 4 files
- Search a minimum of three databasesin 7 of 185, across 6 files
Said here and by no other author read
- Locate the search script before running commands
- Pass natural language queries directly
- Ask the user for their Valyu API key
- Run setup with the provided API key
- Retry the original search after setup
- Pipe results through jq to extract fields
Grouped from the skills themselves: near-identical wordings counted once, and counted by distinct author, so one author publishing three of these counts once. Length counted with cl100k_base; the agent that loads this file may tokenize it differently.