FREE USPS SHIPPING OVER $200Order by 12 PM Pacific Time for same-day shippingTracking on every order

Home / Guides / taxonomy

Types of peptides: what the word guarantees, and how the classes are actually drawn

taxonomyUpdated 2026-08-26Reviewed by Mike Vance, Chief Research OfficerResearch use only
GHK-Cu research vial beside a mass spectrometer and a tray of sealed vials
Short answer

Peptides are classified four ways at once: by chain length, by topology, by chemical modification, and by the receptor family or biological role that groups them functionally. A single molecule sits in all four schemes. Which classification matters depends on whether you are running a separation, reading a certificate, or filing paperwork.

Key facts
  • The peptide-protein boundary is a convention, set near fifty residues in biochemistry teaching and at forty residues in US drug law since 23 March 2020.
  • Every peptide has a position in four independent classification schemes: chain length, topology, chemical modification, and functional or receptor-family grouping.
  • Disulfide isomers and D-for-L substitutions are invisible to mass spectrometry, so mass alone cannot confirm identity for those classes.
  • Production route determines the impurity population, with sequence-related impurities dominating synthetic material and host-derived impurities dominating recombinant material.
  • Purity by RP-HPLC area percent says nothing about water content or counterion load, which are separate determinations and heavier on highly basic sequences.

What the word actually guarantees

A peptide is a polymer of amino acids joined by amide bonds. That is the whole of what the term promises, which is why a single catalogue can list a 226 Da dipeptide and a lipidated 39-residue chain under the same heading without contradicting itself.

Where peptide stops and protein starts is a convention, and there is more than one convention in circulation. Biochemistry teaching usually puts the line near fifty residues, on the reasoning that shorter chains rarely hold a stable independent tertiary fold. United States drug law draws it at forty. Under the FDA rule that took effect on 23 March 2020, an alpha amino acid polymer with a specific defined sequence longer than forty residues is a protein and therefore a biological product; forty or fewer leaves it regulated as a drug. Neither number describes anything chemical. They are lines drawn for different purposes and they disagree by ten residues.

Anyone sorting peptides for practical work ends up holding four schemes at the same time: length, topology, chemical modification, and functional or receptor-family grouping. Every molecule has a coordinate in all four. This reference takes them in that order and then covers what each one tells you before you set up a method or read a certificate.

At a glanceFour axes that classify any peptide
  • Length — predicts synthetic route and impurity count
  • Topology — decides whether mass alone confirms identity
  • Modification — sets hydrophobicity, charge and counterion load
  • Origin — determines which impurities are even possible
  • Functional class — a filing convenience, not evidence

Classification by chain length

Length is the crudest axis and still the most useful first cut, because it predicts how the material was made, how hard it was to purify, and how it will behave in an electrospray source.

Length terms, their usual ranges, and representative molecules
TermResiduesExamplesWhat length implies
Di- and tripeptide2 to 3Carnosine (β-alanyl-L-histidine); glutathione; glycyl-L-histidyl-L-lysineOften crystalline, water soluble, poorly retained on C18 without ion pairing
Oligopeptideroughly 4 to 20Oxytocin (9); somatostatin-14 (14); the pentadecapeptide sequence known as BPC-157 (15)Straightforward solid-phase targets; high crude purity achievable
Polypeptideroughly 20 to 50Semaglutide backbone (31); tirzepatide (39); teriparatide (34)Crude purity falls with length; deletion and insertion impurities accumulate
Small proteinabove 50, or multi-chainInsulin (51 across two chains); most growth factorsUsually recombinant; folding and disulfide pairing become quality attributes

The ranges overlap on purpose. Nobody arbitrates whether a 21-mer is an oligopeptide or a polypeptide, and nothing depends on the answer. What does depend on length is synthetic yield. Each coupling in a solid-phase run is efficient but not perfect, and the fraction of chains that survive every step intact declines multiplicatively. A 15-mer at 99.5 percent per-step efficiency comes off the resin far cleaner than a 39-mer at the same efficiency, and that difference is the reason long sequences carry longer impurity lists on their certificates.

Classification by topology

Topology is about which bonds close the structure. It matters more than length for chromatography, and it decides whether mass alone can confirm identity.

Topological classes and their analytical consequences
TopologyDefining bondExamplesConsequence for analysis
Linear, free terminiBackbone amides onlyMost research oligopeptidesBaseline case; charge states track basic residue count
Linear, terminally modifiedN-acetyl or C-amideAmidated C-termini across many hormone analogsAmidation shifts mass by roughly −0.98 Da against the free acid; easy to miss
Disulfide-bridgedCys–CysOxytocin, vasopressin, somatostatin-14, octreotideDisulfide isomers are isobaric; reduced versus non-reduced comparison is required
Head-to-tail cyclicBackbone amide between terminiCyclic research standards, several fungal peptidesLoss of one water against the linear precursor; different retention behaviour
Side-chain lactam bridgeLys side chain to Asp or GluMelanotan II, bremelanotideConformationally constrained; sharper peaks, altered hydrophobicity
Branched or conjugatedSide-chain acylation or linkerLipidated incretin analogs, DAC-linked GHRH analogsStrong retention, carryover risk, interface activity

The disulfide row is the one that catches people. Two cysteine pairs can close in more than one arrangement, and every arrangement has the same molecular formula. An LC-MS result showing the expected monoisotopic mass says nothing about which pairing you have. Establishing that takes either a reduction and re-run to confirm the bridge count or a proteolytic map with the fragments assigned. Suppliers rarely provide it for small cyclic peptides, and a certificate that quotes only mass and area percent for a two-disulfide sequence is silent on a real attribute.

Classification by chemical modification

Modification is where the taxonomy stops being tidy. Any of these can decorate any backbone, and combinations are normal.

Common modifications and what each one changes
ModificationChemistrySeen inWhat it changes
C-terminal amidationCarboxamide in place of free acidMost native peptide hormones and their analogsCharge at neutral pH; stability to carboxypeptidases
N-terminal acetylationAcetyl cap on the alpha amineAc-SDKP and many research sequencesRemoves a basic site; lowers observed charge states in ESI
Lipidation or acylationFatty acid or fatty diacid via a linkerPalmitoylated and C18/C20 diacid analogsAmphiphilicity, interface adsorption, long RP retention
PEGylationPolyethylene glycol conjugateLonger-acting protein and peptide conjugatesMass distribution rather than a single mass; SEC becomes relevant
GlycosylationO- or N-linked sugarsRecombinant glycopeptidesHeterogeneity that a single purity number cannot describe
PhosphorylationPhosphate on Ser, Thr or TyrKinase substrate peptides+79.97 Da; positional isomers need fragmentation to place
D-amino acid substitutionInverted stereocentreD-Trp in somatostatin analogs, D-residues in GHRPsNo mass change at all; only chiral or comparative methods detect it
Non-coded residuesAib, Nle, Orn, homoarginineAib at positions 2 and 13 in tirzepatideProtease resistance, which is not the same as chemical stability
Metal complexationCoordinated copper or zincCopper-bound glycyl-histidyl-lysineIdentity claim includes the metal and its stoichiometry
Salt and counterion formAcetate, trifluoroacetate, hydrochlorideNearly all synthetic peptidesNet peptide content; two vials of equal mass are not equal in peptide

The last row deserves more attention than it gets. A peptide purified by reversed-phase chromatography with trifluoroacetic acid in the mobile phase comes off as the TFA salt unless it was deliberately exchanged, and TFA can account for a substantial share of the powder mass on a highly basic sequence. The stereochemistry row deserves the same suspicion for the opposite reason: a D-for-L substitution is invisible to mass spectrometry, so a certificate reporting the right mass for a sequence containing designed D-residues has confirmed composition and not configuration.

Classification by origin and production route

Two vials with identical sequences and identical purity figures can carry entirely different impurity populations depending on how the material was made. The route is part of the identity of the lot.

Solid-phase synthesis, run with Fmoc or Boc chemistry, produces sequence-related impurities: deletions where a coupling failed, insertions where a residue double-coupled, capped truncates, incompletely deprotected side chains, and oxidation or racemization picked up during the run. These co-elute close to the target peak because they are nearly the same molecule. USP General Chapter 1503, on quality attributes of synthetic peptide drug substances, sets out these categories and the analytical attributes expected against them; it is written for drug substances but the impurity taxonomy applies to any synthetic peptide.

Recombinant expression produces a different list: host cell protein, residual DNA, misfolded or mispaired forms, N-terminal methionine retention, proteolytic clips from host enzymes. ICH Q6B is the reference frame for how those attributes are specified. A recombinant peptide with a clean RP-HPLC trace can still carry host-derived material that reversed-phase chromatography at 214 nm was never going to see.

Enzymatic hydrolysates sit outside both schemes. Collagen and whey hydrolysates are mixtures defined by average molecular weight distribution rather than by sequence, and the phrase types of peptides is doing something different when applied to them. If a specification lists a molecular weight range instead of a sequence, you are holding a mixture and no identity test on a single mass will characterise it.

Extraction from tissue is now rare for research supply and brings its own set of questions about species origin and adventitious agents. It appears mainly in older literature and in some reference standards.

Functional and receptor-family classes

This is the axis most catalogues sort by, because it maps onto what a research group is looking for. The class labels below are pharmacological descriptions of in-vitro target interaction. They describe what a molecule binds in an assay, and nothing in this table is a statement about outcomes in any organism.

Functional groupings commonly used in research supply, with structural signatures
ClassStructural signatureNamed examplesHandling and analysis note
Incretin receptor ligand analogs30 to 40 residues, non-coded residues, fatty acid or diacid conjugateSemaglutide, tirzepatide, retatrutideInterface-active; aggregation on agitation; long RP gradients
GHRH analogs29 to 44 residues, often N-terminally modified, sometimes linker-conjugatedTesamorelin, CJC-1295, sermorelinLength brings sequence-related impurities; check for the specified linker
Ghrelin receptor agonists and GHRPsShort, 5 to 7 residues, D-amino acids, unnatural aromaticsIpamorelin, GHRP-2, hexarelinMass alone cannot confirm the D-residues; short chains give high crude purity
Melanocortin receptor ligandsCyclic lactam or linear heptapeptide core, His-Phe-Arg-Trp motifMelanotan II, bremelanotide, afamelanotideCyclisation confirmed by the water loss against the linear precursor
Somatostatin analogsDisulfide-bridged octapeptides with D-residuesOctreotide, lanreotide, somatostatin-14Requires disulfide confirmation, not just mass
Antimicrobial peptidesCationic, amphipathic, often helicalLL-37, magainin, polymyxin familyPeak tailing on bare silica; strong adsorption to plastics at low concentration
Cell-penetrating peptidesArginine-rich or amphipathic carriersTAT-derived sequences, penetratin, octaarginineHigh charge density; TFA counterion load can be significant
Matrix and matrikine fragmentsVery short collagen- or elastin-derived motifsGlycyl-histidyl-lysine and its copper complex, palmitoyl tripeptidesPoor C18 retention; metal complexes need stoichiometry stated
Adhesion and targeting motifsMinimal recognition sequences, often cyclisedRGD and cyclic RGD variants, NGRSmall, well-behaved, usually straightforward to characterise
Native peptide hormones and neuropeptidesReference sequences, frequently disulfide-bridged or amidatedOxytocin, vasopressin, substance P, glucagonPharmacopeial monographs exist for several; use them as the identity benchmark
Enzyme substrates and inhibitor peptidesChromogenic or fluorogenic tags on a cleavage motifAMC and pNA substrates, protease inhibitor sequencesDetection wavelength set by the tag, not the peptide bond

One caution about this axis. Functional labels travel further than the evidence behind them, and a class name in a catalogue is a filing convenience. Whether a given molecule does what its class name implies, in which assay system, at what concentration, is a literature question and not a catalogue question.

What the class tells you before you set up a method

Sorting a compound into the schemes above is not academic. It narrows method development before the first injection.

  • Hydrophobicity from the modification axis. A fatty diacid conjugate needs a longer organic gradient and will show carryover on the next blank if the column is not washed properly. A copper tripeptide barely retains on C18 at all and may need ion pairing or a polar-embedded phase.
  • Charge from the sequence. Count the basic residues. That count predicts the dominant electrospray charge states and tells you roughly where to set the scan range before you run anything.
  • Topology decides whether mass is sufficient. Linear and unmodified, mass plus retention time against a reference is a reasonable identity argument. Disulfide-bridged or containing designed D-residues, it is not, and the certificate should say what else was done.
  • Length predicts the impurity profile. Long synthetic sequences carry close-eluting deletion impurities, so resolution near the main peak matters more than total run time.
  • Origin decides what the assay cannot see. Recombinant material needs orthogonal methods for host-derived impurities. Area percent at 214 nm counts peptide bonds and reports nothing about water, counterion, or non-chromophoric residue.

That last point is the most common misreading of a certificate. A purity figure of 99 percent by RP-HPLC is a statement about the chromatographic peak area, and a vial can hold 99 percent pure peptide by that measure while a fifth of the powder mass is trifluoroacetate and water. Net peptide content is a separate determination, and the arithmetic for a prepared solution should be built on it. The vial concentration calculator handles the mg per mL bookkeeping once you know what the vial actually contains.

Reading a certificate against the class

The class determines which tests should be on the certificate. A reasonable minimum, by class:

  1. Any peptide: identity by mass spectrometry with the theoretical and observed masses both printed, purity by RP-HPLC with the gradient, column and detection wavelength stated, and water content by Karl Fischer or loss on drying.
  2. Synthetic peptides: counterion identity and content, plus net peptide content where the material will be weighed for quantitative work.
  3. Disulfide-containing or cyclic: evidence of the correct bridge, whether by reduction comparison or peptide mapping.
  4. Sequences with designed D-residues or non-coded amino acids: something beyond mass, since composition alone cannot distinguish the stereoisomer.
  5. Recombinant material: orthogonal purity, host cell protein and residual DNA where relevant.
  6. Glycosylated or PEGylated: a distribution rather than a single figure, with the method that generated it.

An absent test is not the same as a failed test, and no supplier runs everything. What matters is whether the omissions are stated. Our quality standard page sets out which of these we run by class and which we do not. For handling once a lot is in the freezer, the class-specific reasoning is worked through in the storage and stability guide, and diluent selection is covered in the bacteriostatic water reference.

Regulatory position

Class boundaries carry legal consequences in the United States

The FDA final rule defining the term biological product took effect on 23 March 2020. It fixes the boundary at forty amino acids: an alpha amino acid polymer with a specific defined sequence greater than forty residues is a protein, and therefore a biological product licensed under the Public Health Service Act. Forty or fewer keeps the molecule within the drug pathway. Insulin and several other products transitioned to biologic status on that date.

For synthetic copies of peptides originally made by recombinant means, FDA issued final guidance in 2021 describing an abbreviated new drug application route for certain highly purified synthetic peptides referring to listed drugs of rDNA origin, with impurity characterisation expectations attached. Separately, many native peptide hormones have pharmacopeial monographs that define identity and purity for pharmaceutical grade material.

None of these frameworks apply to research-grade material. A peptide supplied for laboratory use carries a certificate of analysis covering identity and purity for a lot, which is a narrower claim about a different kind of product than any licensed medicine, and it is not a lawful route to human use regardless of which class it falls into.

Status verified 26 August 2026.

FOR LABORATORY AND IN-VITRO RESEARCH USE ONLY. NOT FOR HUMAN OR ANIMAL CONSUMPTION. NOT FOR PERSONAL, MEDICAL, DIAGNOSTIC, THERAPEUTIC, OR RECREATIONAL USE.

Common questions

How many amino acids make something a peptide rather than a protein?

There is no chemical answer, only competing conventions. Biochemistry teaching commonly uses about fifty residues, on the reasoning that shorter chains rarely hold an independent tertiary fold. US drug law uses forty, following the FDA rule effective 23 March 2020 that classifies longer defined-sequence polymers as proteins and therefore biological products. The two lines exist for different purposes and neither describes a property of the molecule.

Which classification should I use when documenting a lot?

All of them, briefly. A useful record states the residue count, the topology including any disulfide or cyclisation, every modification with its position, the production route, and the salt form. That set determines which analytical tests are meaningful and which impurities are plausible. A functional class label such as somatostatin analog is fine as an index term, but it does not substitute for the structural description.

Why does mass spectrometry not always confirm identity?

Because several structural differences leave the mass untouched. A D-amino acid substitution has the same formula as its L counterpart. Disulfide isomers with different pairings are isobaric. Phosphorylation positional isomers differ only in where the group sits. In all three cases the observed mass matches the theoretical value while the structure differs from the one claimed, so confirmation needs chiral analysis, reduction comparison, or fragmentation as appropriate.

What does the counterion have to do with classification?

Highly basic classes carry more of it. Arginine-rich cell-penetrating peptides and cationic antimicrobial peptides hold considerably more trifluoroacetate or acetate per milligram than a neutral short sequence does, because there are more sites to pair with. That makes net peptide content a class-dependent question rather than a formality, and it is the reason two vials of equal labelled mass can differ meaningfully in peptide delivered per milligram weighed.

Are collagen or whey hydrolysates a type of peptide?

They are peptide mixtures, which is a different kind of object. A hydrolysate is specified by molecular weight distribution and source protein rather than by sequence, so no single identity test characterises it and purity as a percentage has no clear meaning. If a specification sheet gives a weight range instead of a sequence and a theoretical mass, treat the material as a mixture and document it that way.

Sources

  • IUPAC-IUB Joint Commission on Biochemical Nomenclature recommendations for amino acid and peptide nomenclature. Establishes symbolism, stereochemical designation and the conventions for naming modified residues used throughout this taxonomy.
  • USP General Chapter <1503>, Quality Attributes of Synthetic Peptide Drug Substances. Sets out the impurity categories characteristic of synthetic peptides and the analytical attributes expected against them, including counterion and net peptide content.
  • ICH Q6B, Specifications for Biotechnological/Biological Products. Frames the attributes specified for recombinantly produced material, supporting the distinction drawn between synthetic and recombinant impurity profiles.
  • FDA final rule defining the term biological product, effective 23 March 2020. Establishes the forty amino acid threshold separating peptides regulated as drugs from proteins regulated as biological products.
FROM THE BENCH

Lot reports, storage data, and what we learn testing them.

A short note when new certificates post, when a stability result surprises us, and when a guide worth reading goes up. No promotions.

Research correspondence only. Unsubscribe in one click. We never sell or share an address.

Shop the compounds

Types of peptides: what the word guarantees, and how the classes are actually drawn research vial, supplied with lot-specific documentationRetatrutideSizes, price per mg and lot certificateTypes of peptides: what the word guarantees, and how the classes are actually drawn research vial, supplied with lot-specific documentationTirzepatideSizes, price per mg and lot certificateTypes of peptides: what the word guarantees, and how the classes are actually drawn research vial, supplied with lot-specific documentationSemaglutideSizes, price per mg and lot certificateTypes of peptides: what the word guarantees, and how the classes are actually drawn research vial, supplied with lot-specific documentationBPC 157Sizes, price per mg and lot certificateTypes of peptides: what the word guarantees, and how the classes are actually drawn research vial, supplied with lot-specific documentationGHK-CuSizes, price per mg and lot certificateTypes of peptides: what the word guarantees, and how the classes are actually drawn research vial, supplied with lot-specific documentationTB-500Sizes, price per mg and lot certificate
Browse all compounds