agentsclimarketplace

Blast search

Skill BioTender-max/awesome-bio-agent-skills/skills/bioclaw/blast-search

Run BLAST sequence similarity searches. Use when the user asks to BLAST a sequence, find similar sequences, identify a gene/protein, or do homology search. Triggers on "blast", "sequence similarity", "homology", "identify sequence".From its SKILL.md

Install
npx -y skills add BioTender-max/awesome-bio-agent-skills --skill blast-search

Assembled from the repository path, not quoted from the project. Check it against their README if it does not work.

One thing to look at

  • no licenseNo license file was found in the repository. Code published without one is not open source by default, so using it at work is a question for whoever answers licensing questions where you are.

SKILL.md

2.2 KB, 580 tokens by cl100k_base, as published. Nobody here has run it

BLAST Search

Run NCBI BLAST+ searches inside the BioClaw container.

When to Use

  • User provides a DNA/RNA/protein sequence and wants to find similar sequences
  • User asks to identify an unknown sequence
  • User wants to check sequence conservation across species

How to Execute

1. Determine BLAST program

InputDatabaseProgram
Nucleotide queryNucleotide DBblastn
Protein queryProtein DBblastp
Nucleotide queryProtein DBblastx
Protein queryNucleotide DBtblastn

2. For local BLAST (sequences provided by user)

# Create query file
cat > /tmp/query.fa << 'EOF'
>query_sequence
ATGCGATCGATCGATCG...
EOF

# Create subject file (if user provides reference)
cat > /tmp/subject.fa << 'EOF'
>reference
ATGCGATCGATCGATCG...
EOF

# Run BLAST
blastn -query /tmp/query.fa -subject /tmp/subject.fa -outfmt 6 -evalue 1e-5

3. For remote BLAST (against NCBI databases)

Use BioPython's NCBIWWW module:

from Bio.Blast import NCBIWWW, NCBIXML
from Bio import SeqIO

# Read sequence
sequence = "ATGCGATCGATCGATCG..."

# Run remote BLAST
result_handle = NCBIWWW.qblast("blastn", "nt", sequence)
blast_records = NCBIXML.parse(result_handle)

for record in blast_records:
    for alignment in record.alignments[:10]:
        print(f"Title: {alignment.title}")
        for hsp in alignment.hsps:
            print(f"  Score: {hsp.score}, E-value: {hsp.expect}")
            print(f"  Identity: {hsp.identities}/{hsp.align_length} ({hsp.identities/hsp.align_length*100:.1f}%)")

4. Output format

Present results in a clear table:

*BLAST Results (top 10 hits)*

• Hit 1: Homo sapiens TP53 gene (98.5% identity, E=1e-45)
• Hit 2: Mus musculus Trp53 gene (89.2% identity, E=1e-38)
...

5. Follow-up suggestions

After showing results, suggest:

  • Multiple sequence alignment of top hits
  • Phylogenetic analysis
  • Domain/motif analysis of the query
  • Structural comparison if protein

What ships with it

Read from the repository

Just SKILL.md. No reference files, no scripts.

Keep looking

Skills are one crate of 325,949. Ordering is by how many stacks a row turns up in, so the top of any crate is what has actually been picked rather than what has the most stars.