Evolution for Everyone — David Sloan Wilson on CBC.

David Sloan Wilson is Distinguished Professor of Biology and Anthropology at Binghamton University in New York and author of Unto Others, Darwin’s Cathedral, and most recently Evolution for Everyone. I confess that I have not yet read the book, though it is near the top of my pile. (Dr. Wilson and I are on the editorial board of Evolution: Education and Outreach, and he was kind enough to have copies sent to all of us).

On Saturday I happened to be listening to the radio and caught an interview with Dr. Wilson on CBC’s Quirks and Quarks program. You can listen here to a discussion of the book and his ideas about evolution, morality, religion, and other subjects. (Another connection: the host, Bob McDonald, received an honourary Doctor of Letters degree from the University of Guelph at the same convocation at which I received my degree).

Enjoy.


Collins and me, again.

Those who read the Wired piece on junk DNA and anti-evolutionism will have noticed that both Francis Collins and I made an appearance (see here for more). It’s pretty clear that our views on genome evolution differ considerably, stemming perhaps from our larger views about how evolution works (see here). It then occurred to me that this is the second time our thoughts on genomes have been juxtaposed recently. The first was in a series of book reviews in Lab Times by Weanée Kimblewood. You can read the reviews here. It seems The Evolution of the Genome was well received; The Language of God, not so much. Of course, Collins’s book ranks #583 on Amazon.com; mine, not so much.

RSS feeds: not just for blogs.

RSS, or Really Simple Syndication, is a system that allows readers to receive “feeds” of new content from various online sources to their computer without having to visit each site individually. By using an aggregator, one can subscribe to his or her favourite blogs, news feeds, and other sources of syndicated information and receive new articles as soon as they are posted online.

An amusing introduction to aggregators is given in this short video by Commoncraft:

[Hat tip: Retrospectacle via Sandwalk]

Generally speaking, subscribing to a blog or science news site is as simple as clicking on the “RSS”, “XML”, “Atom”, or other “Subscribe” link. For example, you will notice to the right some links that allow you to subscribe to the Feedburner feeds for this blog, either by email or to an aggregator, or through the larger DNA Network feed encompassing the posts of multiple genetics blogs. (The only downside for bloggers is that hits do not register if someone reads a post in a feed reader rather than coming to the blog — but hey, they’re reading the post, which is what counts!)

Aggregators can be used for more than just reading news and blogs. Many journals have their tables of contents available as RSS feeds. In addition, you can have scientific literature databases such as PubMed (which is free) and Web of Science (subscription required) automatically perform regular searches and send the results to your aggregator. This latter application is incredibly useful, but it is more complicated than subscribing to blogs or news feeds. Fortunately, the University of California San Diego has been kind enough to provide an illustrated guideline to setting up RSS feeds for these and other literature databases.

A list of client-based and web-based aggregators is available in Wikipedia. Prominent examples include Google Reader, Bloglines, and Netvibes.

Happy reading.


Nonsense from home.

I grew up in and around the small city of Orillia, Ontario. It is a charming place, and was both the hometown of Gordon Lightfoot and the summer home of Stephen Leacock, in the latter case serving as the inspiration for his Sunshine Sketches of a Little Town.

My parents (both since re-married) still live in the area, and I try to get home when I can as a good son should. They also make sure that I am kept up to date with local news of interest, which mostly means stories about the hospital administration’s shenanigans (my mother is a nurse and her husband is an MD, both recently retired and glad of it) and the amazing community project for Zambia that my father and stepmother are hard at work planning.

In addition, my mother enjoys sending me things like the following letter, which appeared in one of the local newspapers. This is, I think, the third rant by a creationist that I have seen in print from this or the smaller paper. A previous one claimed that no one had ever considered the evolution of plants, and thus that creationism must be accurate. I have many botanist colleagues who would be surprised to learn of this omission. In light of the recent poll results that show only 51% of Ontarians accept evolution, I think it is informative. This, by the way, is a lower total than for the USA as a whole, which I find disconcerting. To be fair, the results would probably depend heavily on which part of the province they sampled. Rural areas and small towns differ considerably from the larger cities in various socio-political attitudes. Frankly, I don’t have the time or energy to correct the factual errors and logical fallacies in this latest letter, so I will just post it for your enjoyment (original source).

Evolutionists have their heads in the sand
Orillia Packet & Times
Editorial – Wednesday, May 23, 2007 @ 09:00

Letter to the editor:

Re: M. Brown’s letter “Atheism; a sensible alternative to some”

Atheism: the belief that there is no God. Mr. Albert Einstein, the great theoretical physicist, upon having been asked how much he knew about what there is to be known, said he thought he might know about one hundredth of one per cent without doubt. Most of us know less than he did. All of which leads one to wonder how a person can come to conclude that “there is no God, no creator, no higher intelligence,” when we realize how little we know about anything in general, and origins and beginnings in particular.Atheism does not seem very sensible.

To deny evolution, Mr. Brown writes, is like putting one’s head in the sand, and to dismiss it because we still have apes, is ignorant. Well, where is the evidence for the theory? Where are all the in-between transitional life forms that Mr. Darwin and his followers were sure to be found in the fossil record? After all the digging and searching of the last 150 years, out of hundreds of thousands of fossil discoveries, there is not one clear-cut sample, when there should have been thousands. Dr. Colin Patterson, an evolutionist who was senior paleontologist at the prestigious British Museum of National History and a world-renowned fossil expert, wrote the following about his book entitled “Evolution:” “I fully agree with comments on the lack of direct illustration of evolutionary transitions. If I knew of any, fossil or living, I would certainly have included them. I will lay it on the line; there is not one such fossil for which one could make a watertight argument.”

Surely, this leads open, honest minds to conclude that the theory of evolution and all the resulting evo-babble is unproven, unsubstantiated, without evidence and just plain wrong. As a matter of fact, the earliest fossil record shows that plants and animals appeared suddenly and fully formed, much as we see them today, in complete agreement with the Biblical record. So who, in reality, is putting their head in the sand and ignoring the evidence?

All of which makes one wonder, why evolution is still being taught in our schools, and creation ignored. Why do we so readily allow our children to be misled?

P. Visser

The old saying is not quite accurate: you can go home again, just try not to read the letters to the editor in the local newspaper.


Inter-lineage selection versus "just in case".

I still want to grant the benefit of the doubt to my fellow biologists who recently have made statements about non-coding DNA being potentially useful in the future. Natural selection does not work this way, because it is simply the differential survival and reproduction of entities based on heritable differences. In the most common case, this means individual organisms within populations leaving more or fewer offspring and/or surviving or dying under given conditions in a non-random manner due to heritable trait differences. However, the general principle of natural selection is not restricted to this level, and is a logical consequence in any circumstance in which there is differential survival and reproduction based on inherited variation. There can be selection within the genome among transposons, for example, and some authors also argue that selection can take place among species (as differential speciation and extinction).

The most straightforward way of thinking about natural selection is to imagine that a certain genetic trait is either beneficial or detrimental to an organism, such that it is passed on either more or less commonly to subsequent generations. However, there can be higher-order selection as well, in which some lineages persist longer or branch off to form additional daughter lineages more often than others for non-random reasons. This is not why those traits originated nor why they are maintained from one generation to the next, but it could explain why lineages with those traits are more common or last longer than others.

As an example, consider sex. Sexual reproduction involves the recombination of genes which has two important effects: 1) it allows beneficial mutations to spread more easily in a population, and 2) it prevents the ratchet-like accumulation of deleterious mutations at multiple loci. What this means is that sexual lineages can be expected to evolve more quickly and to last longer than asexual lineages. So, when we look around, we expect to see more sexual lineages than asexual ones, and indeed that is what we see (at least in animals). Sex did not evolve so that lineages would have greater evolutionary potential or would survive for a longer time, but that is nevertheless a significant effect when considering the distribution of biological diversity. However, there is still an issue that sexual reproduction is costly: you only pass on half your genes, you produce “wasteful” males, you have to find a mate, and so on, so we also need to consider immediate benefits that keep the trait around long enough for us to even notice the higher-order effects.

Now back to “junk DNA”. It may be that over the long term, lineages with more non-coding DNA are more flexible and can diverge more often, or that they are more resilient to environmental change and will last longer than those with less DNA. If this is so, then this might explain why we see lineages with lots of non-coding DNA — because those lineages persisted while others disappeared. We would still have to explain the origin of the non-coding DNA and the reason it persists over the shorter term though. There are several possibilities. One, non-coding DNA is beneficial to the organism in some way. Lots of ideas have been proposed for this over the last half century. Two, non-coding DNA could be neutral and is simply not eliminated by selection. Three, non-coding DNA is slightly detrimental, but selection has been too weak (e.g., if populations are small) or mutation too strong (e.g., continual transposable element insertions) for it to be deleted. In any of these situations, it could be possible for non-coding DNA to persist long enough to be co-opted (by chance mutations and subsequent selection) or to have impacts on lineage diversification and/or lifespan.

The problem with this is that species with small genomes are much more common than ones with large genomes and large-genomed species seem to be more sensitive to environmental challenges. So, the most likely scenario is that mutational mechanisms affect DNA amount from the bottom up, while selection comes into play from the top down in terms of effects on cell size and also selection against disruptions of genes. On balance, some lineages end up with large amounts of non-coding DNA, and in some cases this is co-opted into functions like regulation or structure.

It certainly could be that some people are thinking about this from a reasonable perspective based on multiple levels of selection and time scales and are just being sloppy in their descriptions of the net processes. Or maybe they really do think that “junk DNA” is kept because it might become useful. Either way, we need to steer clear of simplified soundbites that obfuscate more than enlighten.


Function, non-function, some function: a brief history of junk DNA.

It is commonly suggested by anti-evolutionists that recent discoveries of function in non-coding DNA support intelligent design and refute “Darwinism”. This misrepresents both the history and the science of this issue. I would like to provide some clarification of both aspects.

When people began estimating genome sizes (amounts of DNA per genome) in the late 1940s and early 1950s, they noticed that this is largely a constant trait within organisms and species. In other words, if you look at nuclei in different tissues within an organism or in different organisms from the same species, the amount of DNA per chromosome set is constant. (There are some interesting exceptions to this, but they were not really known at the time). This observed constancy in DNA amount was taken as evidence that DNA, rather than proteins, is the substance of inheritance.

These early researchers also noted that some “less complex” organisms (e.g., salamanders) possess far more DNA in their nuclei than “more complex” ones (e.g., mammals). This rendered the issue quite complex, because on the one hand DNA was thought to be constant because it’s what genes are made of, and yet the amount of DNA (“C-value”, for “constant”) did not correspond to assumptions about how many genes an organism should have. This (apparently) self-contradictory set of findings became known as the “C-value paradox” in 1971.

This “paradox” was solved with the discovery of non-coding DNA. Because most DNA in eukaryotes does not encode a protein, there is no longer a reason to expect C-value and gene number to be related. Not surprisingly, there was speculation about what role the “extra” DNA might be playing.

In 1972, Susumu Ohno coined the term “junk DNA“. The idea did not come from throwing his hands up and saying “we don’t know what it does so let’s just assume it is useless and call it junk”. He developed the idea based on knowledge about a mechanism by which non-coding DNA accumulates: the duplication and inactivation of genes. “Junk DNA,” as formulated by Ohno, referred to what we now call pseudogenes, which are non-functional from a protein-coding standpoint by definition. Nevertheless, a long list of possible functions for non-coding DNA continued to be proposed in the scientific literature.

In 1979, Gould and Lewontin published their classic “spandrels” paper (Proc. R. Soc. Lond. B 205: 581-598) in which they railed against the apparent tendency of biologists to attribute function to every feature of organisms. In the same vein, Doolittle and Sapienza published a paper in 1980 entitled “Selfish genes, the phenotype paradigm and genome evolution” (Nature 284: 601-603). In it, they argued that there was far too much emphasis on function at the organism level in explanations for the presence of so much non-coding DNA. Instead, they argued, self-replicating sequences (transposable elements) may be there simply because they are good at being there, independent of effects (let alone functions) at the organism level. Many biologists took their point seriously and began thinking about selection at two levels, within the genome and on organismal phenotypes. Meanwhile, functions for non-coding DNA continued to be postulated by other authors.

As the tools of molecular genetics grew increasingly powerful, there was a shift toward close examinations of protein-coding genes in some circles, and something of a divide emerged between researchers interested in particular sequences and others focusing on genome size and other large-scale features. This became apparent when technological advances allowed thoughts of sequencing the entire human genome: a question asked in all seriousness was whether the project should bother with the “junk”.

Of course, there is now a much greater link between genome sequencing and genome size research. For one, you need to know how much DNA is there just to get funding. More importantly, sequence analysis is shedding light on the types of non-coding DNA responsible for the differences in genome size, and non-coding DNA is proving to be at least as interesting as the genic portions.

To summarize,

  • Since the first discussions about DNA amount there have been scientists who argued that most non-coding DNA is functional, others who focused on mechanisms that could lead to more DNA in the absence of function, and yet others who took a position somewhere in the middle. This is still the situation now.
  • Lots of mechanisms are known that can increase the amount of DNA in a genome: gene duplication and pseudogenization, duplicative transposition, replication slippage, unequal crossing-over, aneuploidy, and polyploidy. By themselves, these could lead to increases in DNA content independent of benefits for the organism, or even despite small detrimental impacts, which is why non-function is a reasonable null hypothesis.
  • Evidence currently available suggests that about 5% of the human genome is functional. The least conservative guesses put the possible total at about 20%. The human genome is mid-sized for an animal, which means that most likely a smaller percentage than this is functional in other genomes. None of the discoveries suggest that all (or even more than a minor percentage) of non-coding DNA is functional, and the corollary is that there is indirect evidence that most of it is not.
  • Identification of function is done by evolutionary biologists and genome researchers using an explicit evolutionary framework. One of the best indications of function that we have for non-coding DNA is to find parts of it conserved among species. This suggests that changes to the sequence have been selected against over long stretches of time because those regions play a significant role. Obviously you can not talk about evolutionarily conserved DNA without evolutionary change.
  • Examples of transposable elements acquiring function represent co-option. This is the same phenomenon that is involved in the evolution of complex features like eyes and flagella. In particular, co-option of TEs appears to have happened in the evolution of the vertebrate immune system. Again, this makes no sense in the absence of an evolutionary scenario.
  • Most transposable elements do not appear to be functional at the organism level. In humans, most are inactive molecular fossils. Some are active, however, and can cause all manner of diseases through their insertions. To repeat: some transposons are functional, some are clearly deleterious, and most probably remain more or less neutral.
  • Any suggestions that all non-coding DNA is functional must explain why an onion needs five times more of it than you do. So far, none of the proposed unilateral functions has done this. It therefore remains most reasonable to take a pluralistic approach in which only some non-coding elements are functional for organisms.

I realize that this will have no effect on the arguments made by anti-evolutionists, but I hope it at least clarifies the issue for readers who are interested in the actual science involved and its historical development.


ENCODE links.

The ENCODE paper and related commentaries in Nature (June 14):

List of stories about the ENCODE study:

I think the project is very interesting and important, but as I have said before, one study by itself is rarely revolutionary. ENCODE is adding evidence in favour of a revised understanding of genome function. It, along with many other studies, may require us to re-think a few concepts like “regulatory sequences” or “gene”, but this one paper alone is not engaged in battle against some stubborn establishment that steadfastly refuses to consider new possibilities.


More about ENCODE from Scientific American.

It is probably just coincidence, but two articles for which I gave interviews appeared online today. The first, which I discussed in an earlier post, was online in Wired, One Scientist’s Junk Is a Creationist’s Treasure by Catherine Shaffer. The second appeared in the online edition of Scientific American, The 1 Percent Genome Solution by JR Minkel. Both deal with non-coding DNA, though from rather different perspectives. The first is about creationists invoking the discovery (by evolutionary biologists and other scientists) of (indirect indication of) function in (small sections of) non-coding DNA. The second is about the search for those functions through detailed, rigorous scientific analysis.

I know that science writers have a tough job. And I know that we scientists grumble about a lot of what they generate. But this time I want to do something a little different. I want to give readers some idea of what science writers are faced with when they interview a scientist. This is possible because the interview for Scientific American was conducted by email rather than by phone (which I actually prefer). Have a look at the article, and then see how the interview actually proceeded, and think about the challenge of summarizing my answers, which were admittedly somewhat long-winded (some might say carefully worded so as to avoid confusion and to not overlook important points). Note also the kinds of questions that a writer has to develop.

Here are the pertinent sections from the article:

The consortium found that 5 percent of the studied sequence has been conserved among 23 mammals, suggesting that it plays an important enough role for evolution to preserve while species have evolved. But of all the new ENCODE sequences identified as potentially important, only half fall into the conserved group.

These unconserved sequences may be “bystanders, Birney says”—consequences of the genome’s other functions—that neither help nor hurt cells and may have provided fodder for past evolution.

They could also simply maintain a useful DNA structure or spacing between pieces of DNA regardless of their particular sequence, says genomics researcher T. Ryan Gregory of the University of Guelph in Ontario, who was not part of the consortium.

“The biological insights are mainly incremental at this point,” says genome biologist George Weinstock of the Baylor College of Medicine in Houston, which he says is to be expected of such a pilot study. “This is a ‘community resource’ project, like a genome project, that makes lots of new data available to the community, who then dig into it and mine it for discoveries.”

Gregory says the results, although still cryptic, do hint at new functions and a more complicated genome. “This study shows us how far we are from a comprehensive understanding of the human genome.”

And here are my answers to Minkel’s questions reproduced in full:

How much of what the consortium found is new?

– What is new about this study is the fine focus being applied to the search for functional elements. By way of analogy, this study is like a group of 35 treasure hunters with metal detectors and sifters combing the same 35m of a 3.5km long beach. (In fact, the 35m are broken up into 44 discrete stretches of beach, half of them chosen because they are known to contain lots of interesting objects and the other half selected to include areas with varying properties. The plan is eventually to comb the entire beach this way, but this first pass should be taken more as a proof-of-principle than a conclusive assessment).

– Some of the conclusions reinforce ideas that have already been in the literature for several years, for example that the majority of the human genome is transcribed (see, e.g., Wong et al. 2000; Wong et al. 2001). The identification of non-protein-coding transcripts, particularly in areas where this was not thought to occur, is novel. But, again, this particular study is based on only 1% of the genome and one should exercise caution in extrapolating it to the entire human genome.

– Other ideas, such that chromatin structure is important in regulation, are also not entirely new, but these data provide interesting new evidence for them.

How much of what was identified is likely to be functional?

– 5% of the genome sequence is conserved across mammals, and for about 60% of this (i.e., 3% of the genome) there is additional evidence of function. This includes the protein-coding exons as well as regulatory elements and other functional sequences. So, at this stage, we have increasingly convincing evidence of function for about 3% of the genome, with another 2% likely to fall into this category as it becomes more thoroughly characterized.

– The authors report the presence of sequences that are not conserved but show experimental (in the genomics sense) evidence of function. There need not be constraint on base pair sequences if merely the presence of non-coding DNA would fill the role independent of what that DNA is. For example, if it is simply a matter of physical spacing or structural arrangement, then it may not matter what the actual sequence of bases were. On the other hand, the authors argue that these elements “may serve as a ‘warehouse’ for natural selection, potentially acting as the source of lineage-specific elements and functionally conserved but non-orthologous elements between species”. Of course, this would be an effect, not a function, because natural selection does not have foresight and cannot maintain elements because they may someday be useful. Also, they suggest that these regions are “neutral”, meaning that they are “biochemically active” but “do not confer a selective advantage or disadvantage to the organism”. If they have no fitness effects then they cannot have a function in the usual sense of the term; however, it could be that their absence would be detrimental, in which case there would be convincing evidence of function of some sort.

– A large fraction of the sequences analyzed, both in introns and intergenic regions, appears to be transcribed. However, most of this DNA is not conserved and there is no clear indication of function. It could be that the transcripts themselves play a functional role or that the process of transcription but not the transcripts per se contributes an important effect. It could be that the regions they examined, which were typically gene-dense, included transcribed introns (no surprise) plus longer-than-expected regulatory regions such as promoters near but outside of genes (e.g., Cooper et al. 2007), but that on the whole the long stretches of non-coding DNA in between genes are not actually transcribed. Or, it could be that transcription in the human genome simply is very inefficient. For example, the data in this study suggest that 19% of pseudogenes in their sample are transcribed, even though by definition they cannot encode a protein and are unlikely to play a regulatory role. It also appears that in other groups, e.g., plants (Wong et al. 2000), there is lots of intergenic DNA that is not transcribed, which may indicate that this is a process unique to mammals and is not typical of eukaryotic genomes.

– Looking at a broader scale, we must bear in mind that about half the human genome consists of transposable elements. Some of these clearly do have functions (e.g., in gene regulation), but others persist as disease-causing mutagens. It could be that a large portion of these have taken on functions, but this remains to be shown. We are also left with the question of why a pufferfish would require only 10% as much non-coding DNA as a human whereas an average salamander needs 10 times more than we do. The well known patterns of genome size diversity make it difficult to explain the presence of all non-coding DNA in functional terms, even as there is growing evidence that a significant portion of non-coding DNA is indeed functionally important.

What does this tell us about the genome’s organization and evolution?

– This work follows the growing trend in which simplistic assumptions about genome form and function are being overturned. Previous examples include the assumption that each gene encodes one protein product and the associated expectation that there would be a relatively large number of genes in our genome. This study deals a blow to the notion that the human genome is organized and regulated in a simple way, and further suggests that our definition of “gene” may need to be expanded.

– This study shows us how far we are from a comprehensive understanding of the human genome, but it also provides some of the tools that will be needed to achieve this goal.

– The authors begin their paper with a conclusion (p. 799): “The human genome is an elegant but cryptic store of information.” Elegant, in the scientific sense, means “concise, simple, succinct”. This does not strike me as an accurate descriptor for such a complex, redundant evolutionary patchwork.

– This study reinforces the notion that the genome is a legitimate level of biological organization with its own complex evolutionary history.

You advise caution in extrapolating the results. Do you think it more likely that the study over- or under-represents the amount of complexity or underappreciated function in the genome? Why, or what other biases would you expect?

The concern is that it may over-estimate the level of function in the genome, given that they specifically selected regions rich in well-characterized genes for at least half the dataset. Of course, the objective of the study is to identify functional elements, so an aggressive approach to the question is warranted in that context. However, they probably considered few sequences that were not associated with genes in some way, such as long stretches of short repeats or transposable elements. The study does suggest that regulation is more complex than we thought, shows some evidence of function for some noncoding DNA, and indicates that lots of noncoding DNA is transcribed, but beyond that it hasn’t really clarified these issues — nor should it be expected to, as this was a pilot project only.

You say the study “deals a blow to the notion that the human genome is organized and regulated in a simple way, and further suggests that our definition of “gene” may need to be expanded.” Do most biologists believe the genome is simply organized and regulated? What’s the dominant view?

I would say that, for obvious pragmatic reasons, people assume that a system is simple until it is shown to be otherwise. At first, it was surprising that genome size is decoupled from organismal complexity. Then it was surprising that genes are split into coding exons and noncoding introns. Then it was surprising that half the human genome is transposable elements. Then it was surprising that there are only 25,000 genes. Now it is surprising that a significant portion of the noncoding DNA is transcribed and that gene regulation is not a simple on-off system but involves interactions – perhaps even networks – of coding and noncoding segments. I wouldn’t want to speak for “most biologists”, but I think overall we are coming to appreciate that less has been figured out about genome function than we first thought. And that is what makes the future of genomic science exciting.

And what further evidence would tell us whether we should redefine “gene”? (E.g., would we need to find disease mutations associated with these chimeric transcripts?)

It depends on what you want the term “gene” to represent. In its original definition, it did not specify “protein-coding exons” (because these were not discovered until decades later), and instead referred to a generalized notion of a genetic “determiner” (according to Johansen 1909 “The word gene is completely free from any hypothesis; it expresses only the evident fact that, in any case, many characteristics of the organism are specified in the germ cells by means of special conditions, foundations, and determiners which are present in unique, separate, and thereby independent ways”). After the rise of molecular genetics in the ‘50s, the focus shifted to individual protein-coding sequences (hence, “one gene, one protein”), though this was expanded to include the intron-exon arrangement after it was described by Gilbert in 1978. Now we see that “units of genetic specification”, or what we might want the term “gene” to describe, can include exons, introns (especially as they play a role in alternative splicing to generate several proteins from one “gene”), regulatory regions, promoters, noncoding RNAs, and other elements. Maybe we need a word to mean “an associated unit of protein-coding exons that specifies a particular set of protein products” and one for “all sequences that are involved in generating a particular set of protein products, including coding, regulation, and associated processes”. One of these could be “gene” but we’d need another term to refer to the other. It may be rendered more complex if some regulatory elements affect multiple coding regions (hence discussion regarding relative contributions of cis vs. trans mechanisms). I think it is becoming clear enough that there is more to it than simply transcribing the stretch of DNA and splicing out the introns without linking changes in non-exonic elements to deleterious effects. So, it’s not so much a requirement of more experimental work to identify disease mutations as a conceptual decision about what we want the word to mean based on the more fundamental discoveries about regulation, protein-coding, and non-protein-coding function.

Overall, I think Minkel did a very good job with this piece — especially given the complex issues being discussed and the input offered by several scientists.