ProteinIQ
Sign inStart for free
ProteinIQ

Structure analysis

Protein structure search

Search large protein-structure databases with a query fold, then inspect ranked neighbors, coverage, scores, and alignments.

Open workflow

What is protein structure search?

Protein structure search is the process of querying a structural database with a three-dimensional protein model to find geometrically similar entries. FoldSeek represents local tertiary interactions with a structural alphabet, enabling fast candidate retrieval before detailed score, alignment, and biological review.

Use it to find remote-homolog candidates, fold analogs, related domains, or structural neighbors when sequence search is insufficient. The query may be experimental or predicted, but chain choice, domain boundaries, missing residues, and model confidence affect retrieval.

ProteinIQ runs FoldSeek database search directly against supported collections. Results retain identifiers, alignment statistics, TM-score and LDDT context, coverage, E-values, and files for detailed pairwise and biological review.

When to use protein structure search

  • Structural neighbors are needed. Use it for remote homology, analogs, and fold-family context.
  • A reviewed query model is available. Choose an appropriate chain and domain boundary.
  • Hits can be independently assessed. Add sequence, annotation, and experimental evidence for important candidates.

Benefits of protein structure search

  • Remote relationships can be found. It can reveal similarities missed by sequence search.
  • Large databases can be searched. Retrieval scales to structural collections.
  • Native evidence remains inspectable. Alignments and method-specific scores are returned.

Primary limitations

  • Coverage bounds discovery. Absent database entries cannot be found.
  • Query quality changes ranking. Boundaries and coordinate quality matter.
  • Similarity does not prove function. Functional transfer needs independent evidence.

Protein structure search methods and applications

FoldSeek converts tertiary residue neighborhoods into a structural alphabet and applies sequence-search techniques. Database composition also matters: clustered predicted structures can reveal broad fold neighborhoods, while curated experimental entries may carry stronger ligand, assembly, and functional context.

Structure search supports remote-homology discovery, annotation, fold classification, model-quality investigation, target comparison, and template selection. A match may represent a global fold, shared domain, repeat, or local geometry, so inspect the aligned region before transferring database annotation.

How to run protein structure search online

  1. Prepare the query. Select chain, domain, assembly, and state; review missing or low-confidence regions.
  2. Choose databases. Set collections and thresholds for the needed sensitivity and result volume.
  3. Run FoldSeek. Preserve database version, settings, warnings, and failed inputs.
  4. Review matches. Read rank, E-value, coverage, TM-score, LDDT, alignment, and length differences together.
  5. Validate candidates. Confirm important hits with superposition, sequence evidence, annotations, and experiments.

How to interpret protein structure search results

Evaluate E-value, score, coverage, aligned length, TM-score, LDDT, and sequence identity together. Structural similarity alone does not prove common ancestry or biochemical function; confirm architecture, conserved residues, assembly, ligands, taxonomy, and sequence evidence.

How protein structure search works

FoldSeek performs the core structure search directly; downstream homology and function claims still require independent review.

  1. Prepare the query. Choose the relevant chain, domain, assembly, and conformational state, and review missing or low-confidence regions.
  2. Choose databases. Select structural databases and thresholds that match the intended sensitivity and result volume.
  3. Run FoldSeek. Run FoldSeek while preserving database versions, search settings, warnings, and failed inputs.
  4. Review matches. Inspect rank, E-value, coverage, TM-score, LDDT, residue alignment, and query–target length differences together.
  5. Validate candidates. Confirm important candidates with detailed superposition, sequence evidence, curated annotations, and experiments when the claim requires them.

Inputs and outputs

Check formats before running, then inspect and download the result from every workflow step.

Inputs

Structure-analysis inputs

PDBmmCIFFASTATSV

One experimental or predicted protein structure in PDB or mmCIF format.

Outputs

Reviewable results

PDBCSVTSVJSONFILES

Ranked database hits, identifiers, E-values, coverage, TM-scores, LDDT values, alignments, and downloadable files.

On this page

  • What is protein structure search?
  • Protein structure search methods and applications
  • How to run protein structure search online
  • How to interpret protein structure search results
  • How it works
  • Inputs & outputs

Tools for protein structure search

Use these methods to prepare inputs, run the core analysis, inspect outputs, and validate the evidence described in this workflow.

FoldSeek

FoldSeek

Search structure databases or compare and cluster uploaded protein structures

structure-analysisalignment+3
USAlign

USAlign

Align two protein structures and return TM-scores, RMSD, residue correspondence, and superposed coordinates

structure-analysisalignment+4
PDBFixer

PDBFixer

Repair common coordinate-file issues before structural comparison

structure-analysisquality-validation+3
PDB Download

PDB Download

Retrieve experimental structures from the Protein Data Bank

database-searchstructure-analysis+4
AlphaFold Database Download

AlphaFold Database Download

Retrieve predicted protein structures from the AlphaFold Protein Structure Database

database-searchstructure-analysis+3
PDB to FASTA converter

PDB to FASTA converter

Extract protein sequences from coordinate files for sequence-aware review

format-conversionprotein+2
HMMER

HMMER

Search profile hidden Markov models for independent sequence-level homology evidence

sequence-analysiscomparison+2
MMseqs2

MMseqs2

Search and cluster large protein sequence collections

sequence-analysiscomparison+4
DSSP

DSSP

Assign secondary structure and solvent accessibility from protein coordinates

structure-analysisprotein+1
MolProbity

MolProbity

Check model geometry and steric quality before interpreting structural matches

structure-analysisquality-validation+4
SASA calculator

SASA calculator

Calculate solvent-accessible surface area for matched structures

structure-analysisprotein+1
RMSD calculator

RMSD calculator

Superpose comparison structures on one reference and report RMSD values

structure-analysiscomparison+2

Other structure analysis workflows

Compare related approaches based on the molecular system, available evidence, required inputs, and decision you need to support.

Protein fold recognition

Matches a protein sequence to known structural templates when ordinary sequence similarity is too weak to identify the fold reliably.

Multiple protein structure alignment

Places three or more protein structures into a shared correspondence for conserved-core, family, and evolutionary analysis.

Frequently asked questions

One experimental or predicted protein structure in PDB or mmCIF format.

Ranked database hits, identifiers, E-values, coverage, TM-scores, LDDT values, alignments, and downloadable files.

Confirm accession, model, chain, biological assembly, domain boundaries, residue numbering, missing regions, alternate conformations, and prediction confidence. Repair coordinates only when necessary and retain both the original file and every preparation decision.

Use method-native scores together rather than selecting one universal number. TM-score emphasizes length-normalized global fold similarity, RMSD reports geometric deviation over the aligned atoms, and coverage shows how much of each structure actually corresponds.

No. Similar folds can support different functions, and local similarity can occur without shared global architecture. Review residue-level correspondence, domains, ligands, oligomeric state, taxonomy, sequence evidence, curated annotations, and experiments.

A complete protein structure search project is generally quote-based. Current providers describe fold recognition and protein-structure analysis as customized services covering data review, method selection, modeling or comparison, validation, and interpretation rather than publishing one universal project price.

The cost depends on structure or sequence count, database scope, model preparation, method comparison, manual inspection, figures, annotation, and whether experimental follow-up is included. Open-source FoldSeek, US-align, and FoldMason can remove a software-license fee, but they do not remove expert analysis or compute requirements.

ProteinIQ self-service starts at $29 per month for academic Plus and $99 per month for commercial Pro, with the configured protein structure search workflow estimated in credits before submission. Done-for-you analysis is scoped separately.

Start with a workflow you can inspect and edit

Add your inputs, review the settings, and keep every structure, score, table, and file connected to the step that produced it.

Open workflow
ProteinIQ

© 2026 ProteinIQ

Products

  • Bioinformatics tools
  • Workflows
  • PDB viewer
  • API

Solutions

  • Small molecule
  • RNA discovery
  • Antibody engineering
  • Peptide discovery
  • Enzyme engineering
  • Protein engineering
  • Virtual screening
  • Molecular docking
  • Protein structure prediction
  • RNA structure prediction
  • Protein structure alignment
  • Protein design
  • Sequence alignment
  • Phylogenetic analysis
  • Molecular dynamics simulation

Resources

  • Documentation
  • Blog
  • Guides
  • Datasets
  • Changelog
  • Sitemap

Company

  • About
  • Contact
  • Enterprise
  • Pricing
  • Security
  • Trust center
  • Author
  • Legal
  • Terms
  • Privacy policy

Connect

  • LinkedIn
  • X
  • Discord
  • Pricing