Amino Acids & Peptides
← Back 📋 Q-Bank 🏠 All Units
HIGH YIELD ★★★
Proteins & Enzymes · Unit 1 of 26

Amino Acids & Peptides

TMU Lecture 1 — Lin Yu, Dept of Biochemistry & Molecular Biology Harper's ch. 3 — Amino Acids & Peptides, pp. 15–24 Section I definitions come from here
01

Why the course starts here

Biochemistry begins with amino acids for the same reason a language course begins with the alphabet. Twenty small molecules, joined end to end in different orders, produce every enzyme, every receptor, every antibody and every structural protein in your body. Nothing else in the subject makes sense until you can see why these twenty, and why their differences matter.

And the differences are entirely in one place. Every amino acid in a protein has the same backbone — an α-carbon carrying an amino group, a carboxyl group and a hydrogen. The only thing that varies is the fourth attachment, the R group or side chain. So when you are asked why a particular residue sits in a particular place in a protein, the answer is always about its R group: is it charged, is it greasy, can it hydrogen bond, is it big?

The one idea to carry through the whole of Module A

Structure is chemistry, and chemistry is the side chain. A hydrophobic side chain will be buried in the protein's core away from water; a charged one will face outward or form a salt bridge; cysteine's –SH will look for another cysteine to form a disulfide bond. Every level of protein structure you meet in the next three units is the consequence of twenty side chains seeking their most comfortable environment.

Test yourself
  • What is the only part of an amino acid that varies between the twenty? → The R group (side chain)
  • What four groups attach to the α-carbon? → Amino group, carboxyl group, hydrogen, and the R group
  • Why does the side chain determine protein folding? → It decides whether a residue prefers water or the protein's core
02

The general structure ★★

Every amino acid found in proteins is an L-α-amino acid, and both halves of that phrase are examinable. α means the amino group is attached to the carbon immediately next to the carboxyl group. L refers to the configuration around that carbon — and it is a striking fact of biology that proteins use only the L form, never the D.

The general formula

An α-carbon bearing four attachments: an amino group (–NH₃⁺), a carboxyl group (–COO⁻), a hydrogen, and a variable R group.
Because the four attachments differ, the α-carbon is chiral — with one exception.

⭐ The two exceptions worth marks
Which amino acid is NOT chiral, and why?
Glycine. Its R group is a single hydrogen atom, so the α-carbon carries two identical hydrogens and is therefore not asymmetric. Glycine is the only optically inactive amino acid, and being the smallest it fits where nothing else can — which is why it turns up wherever a peptide chain has to bend sharply.
Harper's ch.3, p.22 · TMU Lecture 1 Slide 14
Which amino acid is not a primary amine?
Proline. Its side chain loops back and bonds to the α-amino nitrogen, making a ring — so proline is a secondary amine, strictly an imino acid. That ring is rigid, which is why proline breaks α-helices and turns up in β-bends.
TMU Lecture 1 Slide 12

Two more points the slides make explicitly. Mammals do contain some free D-amino acids — D-serine and D-aspartate occur in brain tissue — and certain bacterial peptides and antibiotics contain them too. But proteins are built exclusively from L-α-amino acids.

⭐ A discrepancy between your slide and Harper's
The slide says the genetic code specifies 20 amino acids. Harper's mentions a 21st. Which is right?
Both, at different levels of detail. The genetic code directly specifies 20 L-α-amino acids, and that is the number to write if asked. But Harper's devotes a section to selenocysteine — the 21st protein L-α-amino acid, in which a selenium atom replaces the sulfur of cysteine. Humans have around two dozen selenoproteins, including the iodothyronine deiodinases that convert thyroxine (T₄) to T₃. It is co-translationally inserted using a specific tRNA and a recoded UGA codon rather than having a codon of its own.
Answer "20" for the count; mention selenocysteine to show you have read around it.
Harper's ch.3, p.16 · TMU Lecture 1 Slides 11, 15
Test yourself
  • What does 'L-α-amino acid' mean? → Amino group on the carbon next to the carboxyl, in the L configuration
  • Which amino acid is achiral, and why? → Glycine — its R group is a hydrogen, so the α-carbon has two identical substituents
  • Which is a secondary (imino) acid? → Proline — its side chain bonds back to the α-amino nitrogen
  • Name the 21st protein amino acid? → Selenocysteine — selenium replaces the sulfur of cysteine
Cysteine (left) and selenocysteine (right) — selenium replaces the sulfur of the thiol group
Cysteine (left) and selenocysteine (right) — selenium replaces the sulfur of the thiol group
Harper's Illustrated Biochemistry, Figure 3–1, p.18
03

Classifying the twenty

There are several ways to sort the twenty, and the useful one is by what the side chain does in water, because that is what predicts where the residue sits in a folded protein.

ClassAmino acidsBehaviour
Non-polar, aliphaticGlycine · Alanine · Valine · Leucine · Isoleucine · Proline · MethionineHydrophobic — buried in the protein core
AromaticPhenylalanine · Tyrosine · TryptophanBulky, largely hydrophobic; absorb UV at 280 nm
Polar, unchargedSerine · Threonine · Cysteine · Asparagine · GlutamineHydrogen bond; Ser/Thr/Tyr are phosphorylation sites; Cys forms disulfides
Acidic (negative)Aspartate · GlutamateNegatively charged at pH 7 — low pI
Basic (positive)Lysine · Arginine · HistidinePositively charged at pH 7 — high pI; histidine buffers near physiological pH

Two further groupings are worth having. The nutritionally essential amino acids — Harper's counts ten that humans cannot make in quantities adequate for infant growth or adult health — must come from the diet. And a number of amino acids are modified after the protein is made: hydroxylation gives 4-hydroxyproline and 5-hydroxylysine in collagen, and methylation, acetylation, prenylation and phosphorylation all extend what a protein can do.

🩺 Why scurvy is a biochemistry disease

Proline and lysine are hydroxylated after incorporation into procollagen, by enzymes that need vitamin C as a cofactor. Without it, collagen cannot form the hydroxyproline cross-links that give it tensile strength.

The result is scurvy — bleeding gums, loose teeth, poor wound healing, perifollicular haemorrhages. Every one of those signs is connective tissue failing, and the lesion is a post-translational modification of two amino acids.

Test yourself
  • Which three amino acids are aromatic? → Phenylalanine, tyrosine, tryptophan
  • Which are acidic and which basic at pH 7? → Acidic: aspartate, glutamate · Basic: lysine, arginine, histidine
  • How many amino acids are nutritionally essential? → Ten
  • Which two modified residues are found in collagen? → 4-hydroxyproline and 5-hydroxylysine
  • Why does vitamin C deficiency cause scurvy? → It is the cofactor for prolyl and lysyl hydroxylase in collagen synthesis
4-Hydroxyproline and 5-hydroxylysine — modified residues found in collagen
4-Hydroxyproline and 5-hydroxylysine — modified residues found in collagen
Harper's Illustrated Biochemistry, Figure 3–2, p.18
Ultraviolet absorption spectra: tryptophan absorbs most strongly at 280 nm, then tyrosine; phenylalanine barely at all
Ultraviolet absorption spectra: tryptophan absorbs most strongly at 280 nm, then tyrosine; phenylalanine barely at all
Harper's Illustrated Biochemistry, Figure 3–7, p.22
04

Zwitterions ★★★

An amino acid carries two ionisable groups that behave in opposite directions: a carboxyl group that wants to lose a proton, and an amino group that wants to gain one. At the pH of blood, both have done so — the carboxyl is –COO⁻ and the amino is –NH₃⁺. The molecule therefore carries a positive and a negative charge at once, and no net charge overall.

Definition — Zwitterion 3′

A molecule that contains an equal number of positively and negatively charged groups and therefore bears no net charge.
Amino acids exist as zwitterions in blood and most tissues.

Why the uncharged form cannot exist in water

Students often draw an amino acid with a neutral –COOH and a neutral –NH₂ at the same time. Harper's is explicit that this structure cannot exist in aqueous solution, and the reasoning is elegant: at any pH low enough to keep the carboxyl protonated, the amino group would also be protonated; at any pH high enough for a neutral –NH₂, the carboxyl would already be –COO⁻.
The two groups simply do not have overlapping neutral ranges. The uncharged form is used only as a drawing convenience when illustrating reactions that do not involve protons.

Test yourself
  • Define a zwitterion? → A molecule with equal numbers of positive and negative charges and no net charge
  • What form does an amino acid take in blood? → The zwitterion — –COO⁻ and –NH₃⁺
  • Why can the fully uncharged form not exist in water? → The pH ranges for a neutral –COOH and a neutral –NH₂ do not overlap
05

The isoelectric point (pI) ★★★

If the charge on an amino acid depends on pH, there must be one particular pH at which the positives and negatives exactly cancel. That pH is the isoelectric point, and it is one of the most reliably examined definitions in the whole subject.

Definition — Isoelectric pH (pI) 3′ · classic Section I term

The pH at which a molecule bears no net charge and therefore does not move in a direct-current electrical field.
Numerically, it is the pH midway between the pKa values of the ionisations on either side of the isoelectric species.

For an amino acid with only two ionisable groups the calculation is unambiguous. For alanine, the α-carboxyl has pKa 2.35 and the α-amino group pKa 9.69, so the isoelectric point is the average of the two — about 6.02. For an amino acid with a third ionisable group in its side chain, you average the two pKa values that flank the neutral form, which is why acidic amino acids have a low pI and basic ones a high pI.

Amino acidpKa₁ (–COOH)pKa₂ (–NH₃⁺)pI
Alanine2.359.69≈ 6.02 — average of the two
Aspartate (acidic)Low — average of the two carboxyl pKa values
Lysine (basic)High — average of the two amino pKa values
🩺 What pI is actually used for

Because a molecule stops moving in an electric field at its pI, this is the whole basis of electrophoresis and isoelectric focusing — the techniques used to separate serum proteins, to detect abnormal haemoglobins, and to identify the monoclonal band of myeloma on serum protein electrophoresis.

At a pH above its pI a protein is net negative and migrates to the anode; below its pI it is net positive and migrates to the cathode. That single sentence lets you predict the direction of migration in any exam question.

Test yourself
  • Define the isoelectric point? → The pH at which a molecule has no net charge and does not migrate in an electric field
  • How is it calculated? → The pH midway between the pKa values flanking the isoelectric species
  • pI of alanine, and from what? → About 6.0 — the average of pKa 2.35 and 9.69
  • Which way does a protein migrate above its pI? → It is net negative, so toward the anode
Protonic equilibria of aspartic acid — an acidic amino acid, low pI
Protonic equilibria of aspartic acid — an acidic amino acid, low pI
Harper's Illustrated Biochemistry, Figure 3–5, p.21
Protonic equilibria of lysine — a basic amino acid, high pI
Protonic equilibria of lysine — a basic amino acid, high pI
Harper's Illustrated Biochemistry, Figure 3–6, p.21
06

Absorbance at 280 nm

One small physical fact earns marks out of proportion to its size. The aromatic amino acids — tryptophan, tyrosine and phenylalanine — absorb ultraviolet light, and tryptophan absorbs most strongly. Harper's notes that tryptophan makes the major contribution to the ability of most proteins to absorb light around 280 nm.

🩺 How every lab measures protein concentration

Because almost all proteins contain tryptophan and tyrosine, reading the absorbance of a solution at 280 nm gives a fast estimate of protein concentration with no reagents and no destruction of the sample. It is the default measurement on every spectrophotometer and nanodrop in every biochemistry lab — including the ones in your own practicals.

The caveat is also anatomical, in a sense: a protein unusually poor in aromatic residues will read falsely low, which is why colorimetric assays such as the BCA method in your Experiment 2 are used when accuracy matters.

Test yourself
  • Which amino acids absorb UV at 280 nm? → Tryptophan, tyrosine and phenylalanine
  • Which contributes most? → Tryptophan
  • What is this used for? → Rapid estimation of protein concentration by spectrophotometry
07

The peptide bond ★★★

Join the carboxyl group of one amino acid to the amino group of the next, losing a molecule of water, and you have a peptide bond. Repeat it a few hundred times and you have a protein. But the bond has one property that shapes everything above it, and it is the reason protein structure is possible at all.

Definition — Peptide bond 3′

The amide linkage formed between the α-carboxyl group of one amino acid and the α-amino group of the next, with the elimination of a molecule of water.
It has partial double-bond character, so the six atoms of the peptide unit are held rigidly in one plane and rotation about the C–N bond is restricted.

Why partial double-bond character is the key to protein structure

The lone pair on the peptide nitrogen delocalises onto the carbonyl, so the C–N bond behaves as though it were partly a double bond. Two consequences follow, and both are examinable:

1. The peptide unit is planar and rigid, and almost always in the trans configuration.
2. Rotation is therefore possible only about the two other backbone bonds — the N–Cα bond (angle φ, phi) and the Cα–C bond (angle ψ, psi).

That restriction is not a limitation, it is the whole point: a backbone free to rotate anywhere would never fold reproducibly. Because only φ and ψ can turn, the chain has a limited, predictable set of conformations — which is exactly what the α-helix and β-sheet of Unit 3 are.

Two conventions to state when you write about peptides. They are always numbered and read from the N-terminus to the C-terminus — and protein synthesis itself proceeds in that direction, which is a favourite MCQ. And the residue names change ending: in a chain, glycine becomes glycyl, alanine alanyl, with only the C-terminal residue keeping its full name.

Test yourself
  • How is a peptide bond formed? → Between the α-carboxyl of one amino acid and the α-amino of the next, losing water
  • What is its key physical property? → Partial double-bond character — planar, rigid, usually trans
  • Which two backbone bonds can rotate? → N–Cα (phi) and Cα–C (psi)
  • In which direction is a peptide written and synthesised? → N-terminus → C-terminus
Dimensions of a fully extended polypeptide chain — note the planar peptide units and the rotatable φ and ψ bonds
Dimensions of a fully extended polypeptide chain — note the planar peptide units and the rotatable φ and ψ bonds
Harper's Illustrated Biochemistry, Figure 3–9, p.23
08

Revision layer

Everything above was to make it make sense. What follows is the cram layer — and note that four of these are Section I definition terms, worth three marks each.

Definitions from this unit — Section I material

TermDefinition
ZwitterionA molecule containing an equal number of positively and negatively charged groups, therefore bearing no net charge
Isoelectric point (pI)The pH at which a molecule has no net charge and does not migrate in a direct-current electrical field; the pH midway between the pKa values flanking the isoelectric species
Peptide bondThe amide linkage between the α-carboxyl of one amino acid and the α-amino of the next, formed with loss of water; has partial double-bond character
Essential amino acidAn amino acid that humans cannot synthesise in amounts adequate for growth and health, and which must therefore be supplied in the diet — ten in number

The twenty, by side-chain class

ClassMembers
Non-polar aliphaticGly · Ala · Val · Leu · Ile · Pro · Met
AromaticPhe · Tyr · Trp  (UV absorbance at 280 nm)
Polar unchargedSer · Thr · Cys · Asn · Gln
AcidicAsp · Glu  (low pI)
BasicLys · Arg · His  (high pI)

The special cases examiners like

Amino acidWhy it is special
GlycineSmallest; R group is H, so not chiral; fits where nothing else can, found at sharp bends
ProlineSide chain bonds back to the α-nitrogen — a secondary amine; rigid ring breaks α-helices
Cysteine–SH group forms disulfide bonds, the only covalent cross-link in tertiary structure
HistidinepKa near 6 — the only side chain that buffers at physiological pH
TryptophanStrongest 280 nm absorber
SelenocysteineThe 21st protein amino acid; selenium replaces sulfur
Final check — can you do these cold?
  • Define zwitterion, isoelectric point and peptide bond in exam wording
  • Classify all twenty amino acids by side chain
  • Explain why glycine is achiral and why proline breaks helices
  • Explain what partial double-bond character does to the backbone, and name φ and ψ
  • Explain why 280 nm absorbance measures protein, and when it fails