Forum for Science, Industry and Business

Sponsored by:     3M 
Search our Site:

 

Program finds lost genes in nematode genome

11.05.2005


’Good to the last amino acid’


This is C. elegans. Its genome was thought to have been completed until a WUSTL computer scientist applied a computer software program he developed which found scores more genes and predicted the existence of over a thousand additional genes.



A computer scientist at Washington University in St. Louis has applied software that he has developed to the genome of a worm and has found 150 genes that were missed by previous genome analysis methods. Moreover, using the software, he and his colleagues have developed predictions for the existence of a whopping 1, 119 more genes.

Michael Brent, Ph.D., Washington University professor of computer science and engineering, used his unique software, TWINSCAN, on the genome of Caenorhabditis elegans (C. elegans). The genome of another nematode C. briggsae, was also used to determine which parts of the sequence have changed since the nearest common ancestor of the two species. He found first of all that TWINSCAN predicted 60 percent of the genes in the C. elegans genome exactly, right own to the last amino acid.


"This (60 percent) is a new level of accuracy for a complex genome," Brent said. "It’s quite a step up from what you see in the human genome, for instance, where not even a third of the genome can be predicted exactly. The 60 percent is the highest accuracy published for a multicellular organism."

C. elegans is a biological model for animal development and genetics, and is the first animal genome to be sequenced, back in 1998. Nematode researchers rely on a genome annotation database called WormBase. Along with confirmed genes, WormBase includes thousands of predicted genes without evidence from complementary DNA (cDNA) or expressed sequence tags (EST), which help locate genes. These predicted genes are derived by a combination of a program from the previous generation and some curation by human experts. Brent and his colleagues say that the accuracy of WormBase can be improved with the use of TWINSCAN predictions. And Brent predicts that the age of the human genome annotator is passing — the future belongs to computer-driven annotation.

Crossing the tipping point

"We’ve crossed the tipping point with gene prediction where it’s becoming clear that machines can beat human annotators and analysts, on average," he said.

Because of the increasing speed of computers, the TWINSCAN analysis of C. Elegans is able to use more accurate models of intron length than previous analyses. This is important for finding exons, which house the coding machinery of proteins. While getting intron length is helpful for gene annotation, the process is 15 times slower than the typical, less accurate methods. Being able to define intron length has implications for the human genome, which is much larger than C. elegans and has an average intron length of about 4,000 base pairs, compared with an average intron length of a couple hundred base pairs in C. elegans.

Brent and colleagues from the Dana-Farber Cancer Institute and Harvard Medical School published their findings in the April, 2005 issue of Genome Research. Brent’s graduate student, Chaochun Wei, is first author on the paper. The research was supported by grants from NIH, NSF, the National Cancer Institute, the National Human Genome Research Institute, and the National Institute of General Medical Sciences.

Brent has brought his bioinformatics skills to many genomes, including those of mammals, other nematode species and most recently the fungus Cryptococcus neoformans. Brent’s approach to gene prediction stands traditional genome annotation on its head because it starts with a computer analysis of the genome sequence, using that as a hypothesis designing experiments to test the hypothesis. The traditional modus operandi is a data-driven approach that starts with sequencing a random sample of tens of thousands of cDNA clones. Whereas the traditional approach leads to sequencing some genes thousands of times and others not at all, Brent’s approach is to sequence each predicted gene once.

"I’ve been building a case that we should start with predictions," he said. "Each gene sequence is more expensive, but because of the lower redundancy you end up with much better coverage of the genome for the same money."

Chess as metaphor

Brent said that some genome researchers have been reluctant to go towards an automated, hypothesis - driven approach because of a lingering sense that anything that’s been looked at by a human will be more accurate than something produced by a machine.

"But look at the world of chess. Fundamentally, humans are better than machines at chess, but if you get a team of ten people with enough expertise, money, and equipment, and the willingness to work for ten years and burn a lot of computer power, they’ll come up with a machine that can beat the world champion. The same principal applies to developing a machine that can reveal the mysteries of our genes. In this case, the necessary investments have been made, but since there is no sanctioned world championship, it is not yet widely known."

Tony Fitzpatrick | EurekAlert!
Further information:
http://www.wustl.edu

More articles from Life Sciences:

nachricht Climate Impact Research in Hannover: Small Plants against Large Waves
17.08.2018 | Leibniz Universität Hannover

nachricht First transcription atlas of all wheat genes expands prospects for research and cultivation
17.08.2018 | Leibniz-Institut für Pflanzengenetik und Kulturpflanzenforschung

All articles from Life Sciences >>>

The most recent press releases about innovation >>>

Die letzten 5 Focus-News des innovations-reports im Überblick:

Im Focus: Color effects from transparent 3D-printed nanostructures

New design tool automatically creates nanostructure 3D-print templates for user-given colors
Scientists present work at prestigious SIGGRAPH conference

Most of the objects we see are colored by pigments, but using pigments has disadvantages: such colors can fade, industrial pigments are often toxic, and...

Im Focus: Unraveling the nature of 'whistlers' from space in the lab

A new study sheds light on how ultralow frequency radio waves and plasmas interact

Scientists at the University of California, Los Angeles present new research on a curious cosmic phenomenon known as "whistlers" -- very low frequency packets...

Im Focus: New interactive machine learning tool makes car designs more aerodynamic

Scientists develop first tool to use machine learning methods to compute flow around interactively designable 3D objects. Tool will be presented at this year’s prestigious SIGGRAPH conference.

When engineers or designers want to test the aerodynamic properties of the newly designed shape of a car, airplane, or other object, they would normally model...

Im Focus: Robots as 'pump attendants': TU Graz develops robot-controlled rapid charging system for e-vehicles

Researchers from TU Graz and their industry partners have unveiled a world first: the prototype of a robot-controlled, high-speed combined charging system (CCS) for electric vehicles that enables series charging of cars in various parking positions.

Global demand for electric vehicles is forecast to rise sharply: by 2025, the number of new vehicle registrations is expected to reach 25 million per year....

Im Focus: The “TRiC” to folding actin

Proteins must be folded correctly to fulfill their molecular functions in cells. Molecular assistants called chaperones help proteins exploit their inbuilt folding potential and reach the correct three-dimensional structure. Researchers at the Max Planck Institute of Biochemistry (MPIB) have demonstrated that actin, the most abundant protein in higher developed cells, does not have the inbuilt potential to fold and instead requires special assistance to fold into its active state. The chaperone TRiC uses a previously undescribed mechanism to perform actin folding. The study was recently published in the journal Cell.

Actin is the most abundant protein in highly developed cells and has diverse functions in processes like cell stabilization, cell division and muscle...

All Focus news of the innovation-report >>>

Anzeige

Anzeige

VideoLinks
Industry & Economy
Event News

LaserForum 2018 deals with 3D production of components

17.08.2018 | Event News

Within reach of the Universe

08.08.2018 | Event News

A journey through the history of microscopy – new exhibition opens at the MDC

27.07.2018 | Event News

 
Latest News

Smallest transistor worldwide switches current with a single atom in solid electrolyte

17.08.2018 | Physics and Astronomy

Robots as Tools and Partners in Rehabilitation

17.08.2018 | Information Technology

Climate Impact Research in Hannover: Small Plants against Large Waves

17.08.2018 | Life Sciences

VideoLinks
Science & Research
Overview of more VideoLinks >>>