ArticleslgStudy

biology

Protein structural phylogenetics

Protein structural phylogenetics is a biology topic covered in the lgStudy science library. This page brings together a partial reference excerpt, illustrations, worked examples, real-world applications and a short study plan, so you can understand Protein structural phylogenetics rather than just read about it. In short: Protein structural phylogenetics (or Structural phylogenetics) is the branch of molecular evolution that incorporates three dimensional information from protein structure to understand phylogenetic relationships, and translates those evolutionary insights into understanding protein structure and function. Protein structures are robust over long evolutionary time scales compared with amino acid sequence.

Key takeaways

  • Protein structural phylogenetics belongs to biology; place it in that map before memorising details.
  • Learn the definition first, then one example that makes the definition concrete.
  • Connect Protein structural phylogenetics to a quantity you can measure, compute or draw — that is where exam questions come from.
  • Reproduce the core statement of Protein structural phylogenetics from memory before moving on to harder problems.

Reference excerpt

Protein structural phylogenetics (or Structural phylogenetics) is the branch of molecular evolution that incorporates three dimensional information from protein structure to understand phylogenetic relationships, and translates those evolutionary insights into understanding protein structure and function. Protein structures are robust over long evolutionary time scales compared with amino acid sequence. The number of protein sequences that can fold into a given structure is astronomical, with one study estimating even a small protein structure with fewer than 100 amino acids can be attained by a number of sequences that exceeds the Avogadro constant. These properties make structures useful for understanding deep evolutionary relationships, where sequences have become saturated with mutations and share very low levels of similarity.

History Protein structures have been used to explore evolutionary relationships since the 1970s. The approach became popularized in the 1990s and 2000s as techniques in structural biology took off, namely X-ray crystallography, nuclear magnetic resonance, and electron microscopy. Throughout this period, several studies investigated deep evolutionary relationships through the analysis of aligned protein structures, for example, the immunoglobulins aminoacyl-tRNA synthetases, and metallo-β-lactamases. However, the field was still constrained by the limited availability of entries in the Protein Data Bank. In the early 2020s, with the arrival of protein structure prediction methods like AlphaFold2, high quality data became readily available. Although structural predictions are still less accurate than solved structures, this has nevertheless led to three dimensional protein structure becoming increasingly important within the field of phylogenetics. This led to recent insights into the evolution of Flavivirus glycoproteins, fungal virulence factors, and gamete fusion proteins. Despite the abundance of protein structural data, the methodologies to analyze these structures have not kept pace with those used to estimate phylogenies from sequence.

Structural data and alignment Inferring phylogenetic trees from protein structure usually relies on a structural alignment. There are numerous software packages available to perform this task, each with their own strengths and limitations.

Methods for estimating phylogenies from protein structure

Atomic coordinates The simplest methods for building phylogenies from protein structures are based on atomic-level comparisons using measures like RMSD and TM-Score, among others. This distance-based approach is often performed using the neighbor joining algorithm. A key limitation in this approach comes from the inability to quantify statistical uncertainty, such as through bootstrap or posterior clade support. Some have used molecular dynamics simulations to estimate bootstrap support, although this approach is computationally demanding. The more advanced methods are model-based, meaning they describe probabilistic generative processes and can provide a more reliable means of quantifying uncertainty in a maximum likelihood or Bayesian phylogenetic framework. The Challis-Schmidler model describes protein structural drift, over long evolutionary time frames, as an Ornstein–Uhlenbeck process. This Bayesian total-evidence model estimates the sequence and structural alignment all within a single analysis. A key limitation in this method comes from the energetically-unrealistic assumption of independent drift across all positions in the protein. This restriction was later addressed by the Larson-Thorne-Schmidler model.

Structural alphabets Protein structures can also be represented as sequences of characters from a structural alphabet. Typically, there is one character assigned to each amino acid residue in the sequence. This enables structural phylogenies to be built using the same methodologies that are used in sequence phylogenetics, including maximum likelihood and Bayesian inference, as a continuous time Markov process. The earliest efforts involved simple alphabets that describe protein secondary structure and surface accessibility. The 3Di alphabet employed by Foldseek is widely used today. This alphabet consists of twenty characters informed by the protein tertiary structure. While 3Di phylogenetics has become widely applied in recent years, its key limitation comes from the standard phylogenetic assumption of independence between sites, a requirement violated by the concept of the 3Di characters, which are defined by tertiary structure interactions.

See also Molecular phylogenetics Molecular clock

References

Worked examples

Example 1 — a first encounter with Protein structural phylogenetics

Start with the simplest possible case. Write down what Protein structural phylogenetics claims or describes in one sentence, then invent the smallest concrete situation in which that sentence is true. In biology, the smallest case is usually a single object, a single equation or a single measurement. Check that every symbol or term in your sentence has a meaning in that case.

Example 2 — changing one variable

Take the situation from Example 1 and change exactly one quantity: double it, halve it, or set it to zero. Predict what should happen to Protein structural phylogenetics before you calculate. Comparing your prediction with the result is the fastest way to find out whether you understand the idea or only the words.

Example 3 — an exam-style question

Typical questions about Protein structural phylogenetics ask you to (a) state it precisely, (b) apply it to given data, and (c) explain a limitation. Practise writing all three answers in under five minutes; the third part is what separates a full-mark answer from an average one.

Applications of Protein structural phylogenetics

In research
Protein structural phylogenetics appears in biology research whenever the underlying quantities have to be modelled precisely. Papers usually cite it as a starting assumption and then explore where it breaks down.
In technology and industry
Engineering practice reuses Protein structural phylogenetics in design rules, simulations and safety margins. Knowing the idea lets you read a specification sheet and understand why the numbers look the way they do.
In the classroom
Protein structural phylogenetics is common in secondary-school and first-year university syllabi. It links to neighbouring topics Molecular evolution, so understanding it makes those chapters shorter.
In everyday life
Look for Protein structural phylogenetics outside the textbook — in sport, cooking, traffic, electronics or the sky above you. An example you found yourself is remembered far longer than one you were given.

Affiliate

Preply — study more efficiently by working with a personal tutor. 50% off.

How to study Protein structural phylogenetics in 20 minutes

  1. Read the reference excerpt below once, without taking notes.
  2. Close the page and write down what Protein structural phylogenetics means in your own words.
  3. Compare your version with the excerpt and mark what you missed.
  4. Work through the three examples above with pen and paper.
  5. Explain Protein structural phylogenetics out loud to somebody else — or to Teacher Smith in the lgStudy chat.

Frequently asked questions

What is Protein structural phylogenetics in simple terms?

Protein structural phylogenetics (or Structural phylogenetics) is the branch of molecular evolution that incorporates three dimensional information from protein structure to understand phylogenetic relationships, and translates those evolutionary insights into understanding protein structure and fu…

Why does Protein structural phylogenetics matter?

Because it connects several biology ideas at once: it gives you a definition you can apply, a quantity you can calculate, and a way to check whether a result is plausible.

How should I study Protein structural phylogenetics?

Read the excerpt, restate it from memory, then work through the examples and applications listed on this page. The five-step study plan above takes about twenty minutes.

What does this page cover?

It gives you a compact reference excerpt plus original lgStudy explanations, examples, applications and study material on Protein structural phylogenetics.

Tags

  • Molecular evolution

Keep exploring