Bioinformatics Guides & Tutorials
Plain-English guides to DNA, RNA and protein analysis: reverse complement, GC content, primer design, ORFs, codon optimization, restriction digests and FASTQ quality.
Looking for quick lookups instead? Browse the reference tables — codon table, restriction enzymes, IUPAC codes and file formats.
Reverse Complement of DNA, Explained
Compute a reverse complement by hand in two steps, then apply it to design a reverse PCR primer from a real target sequence.
5 min readHow to Calculate GC Content of DNA (Formula)
Calculate GC content correctly, then see a worked primer example revealing an 8°C gap between two Tm formulas, and which to trust for annealing temp.
5 min readPrimer Design Basics: Melting Temperature (Tm) and GC Content
A practical guide to designing PCR primers: target length, GC content, melting temperature formulas (Wallace and salt-adjusted), and matching primer pairs.
6 min readOpen Reading Frames (ORFs) Explained
Scan all six reading frames for start and stop codons, follow a worked ORF example, and see exactly how an ORF differs from a CDS or gene.
6 min readCodon Optimization: How It Works, When to Use It
Optimize a gene for E. coli, human, or yeast: see a worked codon-by-codon example, then check the GC, CAI, and repeat trade-offs before ordering.
6 min readRestriction Enzymes and How to Plan a Digest
How restriction enzymes recognise and cut DNA, the difference between sticky and blunt ends, and how to plan a single or double digest for cloning.
6 min readPhred Quality Scores in FASTQ, Explained
Decode a FASTQ quality string by hand, see why averaging Phred scores overstates accuracy, and learn how N50, L50 and NG50 fit together.
6 min readProtein Molecular Weight and pI Calculator
Calculate protein molecular weight, pI and extinction coefficient from a sequence, and see why two pI tools disagree.
6 min readDNA/RNA ng to pmol Calculator (With Formula)
Convert ng to pmol for DNA or RNA with the formula, molar-mass shortcuts, and a free calculator using exact sequence composition, not just length.
5 min readHow to Find a Motif or Pattern in a DNA Sequence
What sequence motifs are, how IUPAC codes describe degenerate patterns, and how to search a sequence for a motif on both strands with mismatches.
5 min readDNA vs. RNA: Key Differences Explained
Compare DNA and RNA by sugar, base, and strand, then trace a real transcription example showing why T-to-U swaps miss the mechanism.
5 min readIUPAC Nucleotide Ambiguity Codes: Full Table
All 11 IUPAC nucleotide codes (R, Y, S, W, K, M, B, D, H, V, N) with meanings, complements, and a worked degenerate-primer design example.
5 min readHow to Analyze an Unknown DNA Sequence
Characterise an unknown DNA sequence step by step: composition, ORFs, restriction sites and primers — or get all four in one sequence analyzer report.
6 min readPrimer Dimers and Hairpins: How to Avoid Them
Diagnose primer dimers, hairpins and mispriming from gel bands, then apply the ΔG cutoffs (-9 kcal/mol for hairpins, -5 for dimers) that prevent them.
7 min read
How to Design a CRISPR Guide RNA (gRNA)
Design CRISPR guide RNAs using PAM rules for SpCas9, SaCas9 and Cas12a, plus position-weighted on-target and off-target scoring.
6 min readHow to Design Site-Directed Mutagenesis Primers
Design point-mutation, insertion, and deletion primers for QuikChange or Q5-style mutagenesis, with flanking Tm rules and DpnI/KLD cleanup.
6 min readIn Silico PCR: Predict Products Before You Order
Simulate PCR against your template to predict product size and flag every off-target binding site before you order primers.
6 min readHow to Read a Plasmid Map
Learn to read plasmid maps and the GenBank format (LOCUS, FEATURES, ORIGIN) behind them, plus a step-by-step order for checking a new construct.
6 min readPairwise vs. Multiple Sequence Alignment
Compare global (Needleman-Wunsch) and local (Smith-Waterman) alignment, learn affine gap-open vs. gap-extend scoring, and see when you need MSA instead.
6 min read
How to Verify a Clone with Sanger Sequencing
Verify a clone by Sanger sequencing: read the chromatogram, trim low-quality ends, and align the trace to its reference sequence.
6 min readGibson vs Golden Gate vs Restriction Cloning
Compare restriction, Gibson, and Golden Gate cloning on scar and fragment limits, then simulate the assembly and design junction primers before ordering.
6 min readReverse Translation: Protein to DNA Sequence
Back-translate a protein to DNA using most-frequent or degenerate IUPAC codons, then score it with the Codon Adaptation Index.
6 min readNearest-Neighbor Tm and ΔG Explained
Calculate oligo Tm and ΔG from nearest-neighbor stacking energies, and use the same math to flag hairpins and 3'-end dimers before a failed PCR.
6 min readHow to Predict a Restriction Digest Gel
Calculate restriction fragment sizes, then simulate the agarose gel with a ladder using Virtual Gel before you pick a percentage or run one.
6 min readHydrophobicity Plots for Transmembrane Domains
Score a sequence on the Kyte-Doolittle scale to flag transmembrane helices past the classic +1.6-2.0 cutoff, then compare with Hopp-Woods.
6 min readWhy and How to Generate a Random DNA Sequence
How a random DNA sequence generator works, why setting a GC target matters, and how to use a scrambled sequence safely as a negative control or filler.
6 min readHow to Fetch a Sequence by Accession Number
Fetch an exact GenBank, RefSeq or UniProt record by accession number, decode NM_/XP_ prefixes, and pin the version suffix for reproducible results.
6 min readConverting Between FASTA and GenBank Files
Convert between FASTA, GenBank and TSV, and pull a CDS or protein straight from GenBank's FEATURES table, complement strand included.
6 min readIn Silico Peptide Mass Fingerprinting Guide
Predict trypsin, Lys-C and chymotrypsin cut sites and peptide masses, including the K/R-proline exception and missed cleavages for mass spec matching.
6 min readBatch Processing Multiple Sequences at Once
Run one operation across a multi-FASTA or chain tools into a pipeline, then export a single CSV table per plate, primer batch, or Sanger read set.
6 min readIdentify Unknown Sequence: DNA, RNA or Protein
Check whether a sequence is DNA, RNA, or protein by its letters, then send it straight to BLAST — one tool does both steps from a single paste.
6 min readCloning a Gene From Scratch: A Complete Workflow From Sequence to Verified Construct
A step-by-step molecular cloning checklist covering source sequence, assembly strategy, primer design, in silico assembly, construct QC, and Sanger verification.
8 min readCRISPR Knockout Workflow: gRNA to Verified Edit
Rank guide RNAs, design genotyping primers before you edit, screen clones by PCR, and confirm the edit by sequencing in the right order.
8 min readProtein Expression Construct Checklist
Follow a 5-step checklist that catches cryptic ribosome-binding-site errors codon optimizers miss, covering codon usage through construct QC.
7 min read
Why Isn't My PCR Working? Troubleshooting Guide
Diagnose no product, multiple bands, primer dimers, and smears in the order they usually turn out to be the real cause, with a fix for each.
7 min read
Restriction Digest Not Cutting or Wrong Bands
Troubleshoot no bands, extra bands, or an odd-sized uncut lane by symptom, covering star activity, partial digestion, and topology mismatches.
7 min readColony PCR Troubleshooting: No Bands, Wrong Size
Diagnose colony PCR by symptom: no product, faint bands, or inconsistent colonies. Learn junction-spanning primer design for unambiguous results.
7 min readWhy Is My Codon-Optimized Gene Expressing Poorly? Common Mistakes to Check
Explains why a codon-optimized gene can still express poorly, covering GC extremes, mRNA structure, cryptic motifs, and protein-level solubility limits.
7 min readWallace vs. Salt-Adjusted vs. Nearest-Neighbor Tm: How Much Do These Formulas Actually Disagree?
A data-driven comparison of three primer Tm formulas across primer length and GC content, showing exactly how many degrees C they disagree by and why.
9 min readGolden Gate Overhang Fidelity, Scored Against Real Ligation Data: Three Published Sets vs. Two Naive Designs
Real weakest-link fidelity scores for three published Golden Gate/MoClo overhang sets and two naive designs, computed from T4 and BsaI-HFv2 ligation-count data.
11 min readType IIS Enzymes for Golden Gate and MoClo Assembly: BsaI, BsmBI, BbsI, and SapI Compared
Verified recognition sites, cut sites, overhang lengths, and reaction temperatures for BsaI, BsmBI, BbsI, and SapI, mapped to MoClo, the plant Golden Gate toolbox, CIDAR MoClo, and GoldenBraid.
13 min readThe Anderson Promoter Collection: A Verified Reference Table for BBa_J23100-Series Constitutive Promoters
Verified reference table of the 20 Anderson (BBa_J23100-series) constitutive promoters from the iGEM Registry, with sequences cross-checked against an independent mirror and honestly labeled relative-strength confidence.
11 min readWiring an AI Agent to Real Bioinformatics Tools via MCP
A worked walkthrough of calling SeqBench's MCP server from an AI agent: one JSON-RPC tools/call example and a full Golden Gate domestication chain with gate/provenance data.
11 min readSnapGene vs Benchling vs ApE vs SMS vs SeqBench: A 2026 Cloning Software Comparison
An axis-by-axis comparison of SnapGene, Benchling, ApE, SMS, and SeqBench for cloning design: pricing, assembly simulation, APIs, and overhang fidelity checking.
12 min read
Free Sequence Tools for an iGEM Season: A Phase-by-Phase Checklist
A phase-by-phase checklist mapping free sequence tools to iGEM season tasks: primers, Golden Gate overhang fidelity, pre-synthesis QC, colony verification.
12 min readThe Trust Problem With AI-Designed DNA Constructs
Why AI assistants for cloning and CRISPR design tend to report unverified constructs as finished, and why the fix has to live in code, not in a prompt.
9 min read