DNA Protein Translator and ORF Finder

Find open reading frames across all six frames, translate DNA or RNA with the NCBI Standard Code, and compute oligo melting temperature and GC content.

Mode

Paste a coding DNA or RNA sequence. FASTA header lines and whitespace are stripped.

Sequence alphabet
Protein code

This version uses the NCBI Standard Code. Variant codes are documented.

All three forward frames and all three reverse-complement frames are always shown.

GC content

Paste a sequence

Standard genetic code, all six reading frames.

Unambiguous bases
0
G plus C bases
0
Ambiguous bases
0

Ambiguous bases are excluded from the GC denominator.

Cleaned sequence

Enter a sequence.

Reverse complement

Enter a sequence.
Six reading frames
FrameOffsetCodonsProteinNotes
Selected frame detail: +1
CodonTripletAmino acid
Export

six-frame translation, ORF scan, Wallace and nearest-neighbor Tm, GC content How?

How this is calculated

The input is uppercased, FASTA header lines and whitespace are removed, and RNA U bases are converted to DNA T bases. Auto mode rejects a mixed T and U alphabet because the sequence cannot be classified cleanly.

Translate mode uses the NCBI Standard Code, transl_table 1. Each full codon is translated in the three forward frames and the three reverse-complement frames. Ambiguous codons are translated as X, trailing incomplete codons are omitted and reported, and stop codons remain visible as asterisks.

The ORF finder scans all six frames, opening at a start codon (ATG by default, optionally GTG or TTG) and closing at the first in-frame stop, reporting the longest open reading frame per stop. Length includes the stop codon; coordinates are 1-based on the plus strand, and a reverse-strand ORF reads start greater than end. The minimum length defaults to 75 nt and is adjustable.

Melting temperature uses the Wallace 2 plus 4 rule below 14 nt and the SantaLucia 1998 nearest-neighbor model at 14 nt and up, with adjustable sodium and oligo concentration and a stated salt correction. The active model and its valid length range are shown. Tm mode takes an unambiguous DNA A, C, G, and T oligo; RNA duplex melting uses different nearest-neighbor parameters, so RNA is rejected. GC content is reported as G plus C over the unambiguous bases, available on its own as GC content mode.

Formula: six-frame translation, ORF scan, Wallace and nearest-neighbor Tm, GC content

Sources

  1. NCBI Genetic Codes. National Center for Biotechnology Information. Retrieved .
  2. NCBI ORFfinder. National Center for Biotechnology Information. Retrieved .
  3. SantaLucia 1998, PNAS 95:1460 (unified nearest-neighbor parameters). Proceedings of the National Academy of Sciences. Retrieved .
  4. Wallace et al. 1979, Nucleic Acids Research 6:3543. Nucleic Acids Research. Retrieved .

Method last reviewed