From owner-celegans@net.bio.net Thu Nov 03 22:00:00 1994 Path: biosci!agate!howland.reston.ans.net!pipex!sunic!trane.uninett.no!daresbury!not-for-mail From: Newsgroups: bionet.celegans Subject: Re: Percent coding? Date: 4 Nov 1994 16:34:00 -0000 Lines: 15 Sender: lpddist@mserv1.dl.ac.uk Distribution: bionet Message-ID: <39dnpo$7k@mserv1.dl.ac.uk> Original-To: celegans@dl.ac.uk Tim Lindblom (lindblom@zookeeper.zoo.uga.edu) asks about current percentages of predicted coding vs noncoding regions in C elegans. Here is some information from the genome sequencing project. It appears to vary with region, not surprisingly, and you should remember that we are aiming to sequence gene rich regions. For one 1.7Mb region 31% of the sequence is coding (i.e. in exon). We have another 400kb region that is about 37% coding. I don't have a figure for everything sequenced so far, but expect that it would be biased because the region we are sequencing appears to be relatively gene dense. I would predict an average value of well under 30% for the whole genome. Richard .