<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD 2.3 20070202//EN" "http://dtd.nlm.nih.gov/publishing/2.3/journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
	<front>
		<journal-meta>
			<journal-id journal-id-type="nlm-ta">J Comput Sci Syst Biol</journal-id>
			<journal-id journal-id-type="publisher-id">opg</journal-id>						
			<journal-title>Journal of Computer Science &amp; Systems Biology</journal-title>			 
			<issn pub-type="epub">0974-7230</issn>
			<publisher>
				<publisher-name>OMICS Publishing Group</publisher-name>
				<publisher-loc>India, USA</publisher-loc>
			</publisher>
		</journal-meta>
		<article-meta>	
			<article-id pub-id-type="doi">10.4172/jcsb.1000033</article-id>		
			<article-id pub-id-type="publisher-id">000063</article-id>
			<article-categories>
				<subj-group subj-group-type="heading">
					<subject>Research Article</subject>
				</subj-group>
				<subj-group subj-group-type="Discipline">
					<subject>Biochemistry</subject>
				</subj-group>
				<subj-group subj-group-type="System Taxonomy">
					<subject>Proteomics</subject>
					<subject>Bioinformatics</subject>
					<subject>Genomics</subject>
					<subject>Transcriptomics</subject>
					<subject>Biomarkers</subject>
				</subj-group>
			</article-categories>
			<title-group>
				<article-title>Damages in Metabolic Pathways: Computational Approaches</article-title>
			</title-group>
			<contrib-group>
				<contrib contrib-type="author">
					<name>
						<surname>Lakshminarasimhan</surname>
						<given-names>Nruthya</given-names>						
					</name>																			
				</contrib>								
				<contrib contrib-type="author">
					<name>
						<surname>Tagore</surname>
						<given-names>Somnath</given-names>
					</name>								
						<xref ref-type="corresp" rid="cor1">&ast;</xref>											
				</contrib>									
			</contrib-group>
			<aff>Department of Biotechnology &amp; Bioinformatics, Dr D Y Patil University, Plot 50, Sec 15, CBD Belapur, Navi Mumbai 400614, India</aff>						
			<author-notes>
				<corresp id="cor1">&ast; To whom correspondence should be addressed: Somnath Tagore, Department of Biotechnology &amp; Bioinformatics, Dr D Y Patil University, Plot 50, Sec 15, CBD Belapur, Navi Mumbai 400614, India; E-mail: <email>somnathtagore@yahoo.co.in</email></corresp>
			</author-notes>
			<pub-date pub-type="collection">
			     <month>06</month>
				 <year>2009</year>
			</pub-date>
			<pub-date pub-type="epub">
				<day>24</day>
				<month>06</month>
				<year>2009</year>
			</pub-date>			
			<volume>2</volume>
			<issue>3</issue>
			<fpage>208</fpage>
			<lpage>215</lpage>
			<history>
			<date date-type="received">
			     <day>15</day>
				 <month>05</month>
				 <year>2008</year>
			</date>
			<date date-type="accepted">
			      <day>23</day>
				  <month>06</month>
				  <year>2009</year>
			</date>
			</history>
			<permissions>			 
			<copyright-statement><bold>Copyright:</bold> &copy; 2009 Lakshminarasimhan N, et al.</copyright-statement>
			<copyright-year>2009</copyright-year>
			<license license-type="open access">
			 <p>This is an open-access article distributed under the terms of the Creative Commons Attribution License, which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited.</p>
			 </license>
			 </permissions>			
			<abstract>
				<p>Damage analysis in metabolic pathways is one of the most highlighted fields in systems biology area. Metabolic
errors are studied with respect to inherited enzyme deficiencies. The damages usually occur due to changes in a particular enzymatic reaction or deletion of a step. Some enzyme disorders can also cause the particular reaction not to occur causing accumulation or reduced presence of a metabolite. This leads to dysfunction and disease state. Furthermore, the deletion of a path may change the end product of a reaction and can give rise to disorders. Thus, if the pathway leading to these changes can be determined, it would be easier to determine the disorders. Various statistical methods have been employed to determine the damages in the metabolic pathway. One such method is graph analysis technique. Here a graph is represented in the form of nodes and edges and the various properties may be determined. The wet lab results may be interpreted using statistics. Other studies of metabolomics can be performed using flux based analysis, fourier -transformation, Lp and power laws.</p>
			</abstract>
			 <kwd-group>
				<kwd>Enzymatic changes</kwd>
				<kwd>Enzyme deficiency</kwd>
				<kwd>Graph analysis</kwd>
				<kwd>Metabolic blocks</kwd>
				<kwd>Metabolic errors</kwd>										
			</kwd-group>
			<custom-meta-wrap>
				<custom-meta>
					<meta-name>citation</meta-name>
					<meta-value>Lakshminarasimhan N, Tagore S (2009) Damages in Metabolic Pathways: Computational Approaches. J Comput Sci Syst Biol 2: 208-215. doi:<ext-link ext-link-type="doi" xlink:href="10.4172/jcsb.1000033">10.4172/jcsb.1000033</ext-link></meta-value>
				</custom-meta>
			</custom-meta-wrap>
		</article-meta>
	</front>	
	<body>
	     <sec>
		     <title>Introduction</title>
			 <p>Systems biology is an emerging multi-disciplinary field in which the behavior of complex biological systems is studied by considering the interaction of all cellular and molecular constituents rather than using a reductionist approach (Kirschner MW, 2005; Liu ET, 2005). As a part of systems biology, we would concentrate on damages in metabolic pathways. Damage in the metabolic pathway is an occurrence of irregularity of the pathway which causes a particular physiological change. Garrod et al., 1902 was the first to propose the existence of metabolic errors, based primarily on his studies of alkaptonuria. The chemical conversions in the body may occurs as sequential chemical steps, called a metabolic pathway. For example, the breakdown of tyrosine in the body involves a number of discrete chemical steps. Garrod&rsquo;s studies indicated that only one step in this metabolic pathway is defective, leading to accumulation of substances that cannot get past this block. Enzymes are also responsible for the catalysis of the biochemical reactions in metabolic pathways (Herzyk et al, 2004; Hwang et al, 2007). Analogous enzymes are able to catalyze the same reactions, but they present no significant sequence similarity at the primary level, and possibly different tertiary structures as well (Miranda et al, 2008).</p> 
			 <p>Some enzyme alterations do not allow the particular reaction to take place, which causes accumulation or reduced
presence of a metabolite which leads to dysfunction and disease state. The deletion of a path may change the end product which would cause disorders (Guo Z, 1986; Stuart et al, 2008). Damages can be caused by genetic alterations and mutations reduced consumption of metabolite or excessive presence of them. With the subsequent discovery of the genetic role of DNA and its structure as well as greater understanding of protein structure, the genetic bases for inherited changes in protein function were clarified
(Kelekar et al, 2007; Kravchuk et al, 2007; Doetsch et al, 2004). Furthermore, changes in the structure of a gene are likely to change the structure of the enzyme (Speed et al, 2002; Cordwell SJ, 1999). Most random changes are likely to make the enzyme function less well, and many mutations completely abolish the activity of the enzyme (Speed et al, 2002; Cordwell SJ, 1999; Graber et al, 2009). Also, some activities are recessive, since only one functional copy of the gene is often sufficient to carry out the required enzyme activity (Speed et al, 2002; Herzyk et al, 2004; Lorch Y, 1999). Similarly, many proteins consist of two or more distinct polypeptide subunits, each coded by a different gene. Mutations in any one of the genes may interfere with function
of the enzyme, producing similar or indistinguishable phenotypes (Matschinsky, FM, 1996; Cordwell SJ, 1999).</p>
			<p>Using graph analysis attempts have been made to study metabolic functions . (Herzyk et al, 2004; Valiente et al, 2004). The biological damages are generated in a metabolic network to experimentally determine the existence of an organism with damage. The enzyme might be altered in order to sum the importance of the enzyme in a particular metabolism (Schmitz-Peiffer C, 2000). It was found that enzyme associated with high damage are involved in the production of compounds of small connectivity that connect important parts of the metabolism on the other hand, highly connected compounds tend to be redundant since they are
produced by many reactions (Herzyk et al, 2004; Schmitz- Peiffer C, 2000; Stuart et al, 2008). Using a graph analysis of its metabolism, the extent of the topological damage generated in the metabolic network by the deletion of an enzyme has been related to the experimentally determined viability of the organism in the absence of that enzyme. It is shown that the network is robust and that the extent of the damage relates to enzyme importance. It is also predicted that a large fraction (91%) of enzymes causes little damage when removed, while a small group (9%) can cause serious damage. Experimental results confirm that this group contains the majority of essential enzymes. The results may reveal a universal property of metabolic networks.( Lemke
N, Her&eacute;dia F et al,2004).</p> 		 
	 </sec>
	 <sec>
	 	<title>Types of Damages</title>
			<sec>
				<title>Gene Expression Level</title>
					<p>These include changes at the expression levels. Some genes are differentially expressed giving rise to faulty results. A very small change in the expression of a particular gene may have dramatic physiological consequences if the protein encoded by this gene plays a catalytic role in a specific cell function (Dudoit et al, 2002; Herzyk P, 2004; Matschinsky FM, 1996; Krogh et al, 1996). From the biological prospective, even a very small change in the expression of a particular gene may have dramatic physiological consequences if the protein encoded by this gene plays a catalytic role in a specific cell function (Dudoit et al, 2002; Herzyk P, 2004; Matschinsky FM, 1996; Schmitz-Peiffer C, 2000). Many other downstream genes may amplify the 
signal produced by a particular gene, increasing their chance of being selected by certain gene selection protocols. But, for a regulatory gene, however, the chance of being selected by such methods may diminish as one keeps hunting for downstream genes that tend to show much bigger changes in their expression (Speed et al, 2002; Herzyk et al, 2004; Gholson et al,1982; Graber et al, 2009). As a result, the initial list of candidate genes may be enriched with many effector genes that do little to elucidate more fundamental mechanisms of biological processes. Furthermore, the characteristic of the regulatory genes is that their gene expression changes may be low, but they are highly correlated with the downstream highly expressed genes (Herzyk et al,
2004; Brown et al, 1997).</p> 
			</sec>
			<sec>
				<title>Mutations in the Gene</title>
					<p>Mutations in the gene may be considered similar to gene expression changes, rather it can be considered as reason for differential gene expression. Genetic mutations cause changes in the gene that code for a particular protein participating in the pathway. If the gene participating in the functioning of the enzyme is mutated its role will change. The enzyme responsible for performing a particular catalytic reaction would fail, resulting in metabolite accumulation, causing disorders. (Cordwell SJ, 1999; Doetsch et al, 2004). For example, autophagy is involved in cellular development and differentiation and may have a protective role against aging. It is also a form of cell death when allowed to proceed to completion and when cells unable to undergo apoptosis are triggered to die. It is often unclear whether it is directly involved in initiation and/or execution of cell death, or if it merely represents a failed or exhausted attempt to preserve cell viability (Abedin et al, 2007; Karantza- Wadsworth et al, 2007). Recent studies indicate that autophagy may play an active role in programmed cell death,
but the conditions under which autophagy promotes cell death versus cell survival remain to be resolved. Defective autophagy
has been implicated in tumorigenesis, as the essential autophagy regulator <italic>beclin1</italic> is monoallelically deleted in human breast, ovarian, and prostate cancers. <italic>beclin1</italic> is the mammalian ortholog of the yeast <italic>atg6/vps30</italic> gene, which is required for autophagosome formation. Beclin1 expression in human MCF7 breast cancer cells suppresses tumorigenesis. Recent studies revealed that autophagy enables tumor cell survival in vitro and in vivo when apoptosis is
inactivated, as commonly occurs in human cancers (Kelekar et al, 2007; Kravchuk et al, 2007).There are many such examples which can explain the gene mutation to be a reason for metabolic damage .</p>
			</sec>
			<sec>
				<title>DNA Structure Alteration</title>
					<p>Compounds that form chemical bonds with DNA represent several important classes of anticancer drugs (Doetsch
et al, 2004). Increased removal or replicative bypass of DNA adducts have been found to be among the major mechanisms by which cancer cells reverse the effects of intraand inter strand cross-linking drugs (Lorch et al, 1999; Doetsch et al, 2004). Resistance of cancer cells to chemotherapeutic drugs is a major problem encountered in the treatment of tumors. The basis for resistance has been studied extensively and can be drug-specific, or nonspecific as in the case of alkylating and platinum drugs (Levine et al,
1999; Albertson DG, 2006; Lukas et al, 2005; Doetsch et al, 2004). Despite the existence of a large body of data characterizing DNA cross-link repair, the specific components of DNA damage processing pathways and their individual contributions to the repair of cross-links are poorly understood (Lorch et al, 1999; Kelekar et al, 2007; Doetsch et al, 2004). Exploring the influence of these pathways on cell survival and defining their potential relationships with each other are crucial for understanding the modes of action of DNA cross-linking anticancer drugs used currently in the clinic (de Andrade, 2006; Lang J, 1999; Levine B, 1999;
Albertson DG, 2006; Lukas et al, 2005).</p>
			</sec>
			<sec>
				<title>Mitochondrial Damages</title>
					<p>Mitochondria primarily generate ATP, promoting the closure of KATP-channel (ATP-sensitive K+ channels) and,
as a consequence, depolarization of the plasma membrane (de Andrade et al, 2006). Mitochondrial defects in &beta;-cells perturb glucose stimulated insulin secretion and might be associated with diabetes (Matschinsky FM, 1996). At the clinical level, one of the most direct indications that mitochondrial dysfunction could impair insulin release comes from the association between mutated mitochondrial genome and maternally inherited diabetes (Rorsman P, 1997; Lang J 1999).</p>
			</sec>
			<sec>
				<title>Transcription Level Changes</title>
					<p>The damages related to transcription may be a result of the failure to detect the transcription binding sites or the wrong signal from the TF for the detection of upstream/ downstream regions (Caisheng et al, 2009). For example,
let us consider a study on the physiological behaviour of <italic>Lactococcus lactis</italic> subsp. <italic>cremoris</italic> MG 1363 was characterized in continuous culture under various acidic conditions (pH 4&middot;7&ndash;6&middot;6) was performed. Biomass yield was diminished in cultures with low pH and the energy dedicated to maintenance increased due to organic acid inhibition and cytoplasmic acidification. Under such acidic conditions, the specific rate of glucose consumption by the bacterium increased, thereby enhancing energy supply. This acceleration of glycolysis was regulated by both an increase in the concentrations of glycolytic enzymes (hierarchical regulation) and the specific modulation of enzyme activities (metabolic regulation). However, when the inhibitory effect of intracellular pH on enzyme activity was taken into account in the model of regulation, metabolite regulation was shown to be the dominant factor controlling pathway flux. The changes in glycolytic enzyme concentrations were not correlated directly to modifications in transcript concentrations. Analyses of the relative contribution of the phenomena controlling enzyme synthesis indicated that translational regulation had a major influence compared to transcriptional regulation. An increase in the translation efficiency was accompanied by an important decrease of total cellular RNA concentrations, confirming that the translation apparatus of <italic>L. lactis</italic> was optimized under acid stress conditions.</p>
<p>Transcriptional damage may be caused due to the chromosomal changes and may lead to protein misfolding. This leads to protein structure damage and may not result in the formation of the particular protein. The transcriptional changes also lead to protein frameshift which would result in wrong metabolite metabolite formaition.</p>
			</sec>
			<sec>
				<title>Enzymatic Dysfunction</title>
					<p>Enzymes are proteins that catalyze specific metabolic steps. Many enzymes catalyze a single step; most catalyze only a few. Enzymes are highly specific with respect to the substrates on which they act (Herzyk P et al, 2004; Cordwell
SJ, 1999). Many metabolic pathways consist of a series of chemical reactions, each catalyzed by a separate enzyme. It is very necessary to detect those enzymes which may be deleted from the path because the metabolite to be formed in the reaction catalyzed by that particular enzyme may not be formed (Herzyk et al, 2004; Gholson et al., 1982; Guo Z, 1986). This blocks the further process. Sometimes low availability of the enzyme may cause the formation of an altogether new product (Graber et al, 2009). If they occurs enzymatic alterations the particular end compound would not form which may lead to major changes in the metabolism
process depending on its occurrence in the pathway (Herzyk et al, 2004; Gholson et al, 1982; Schmitz-Peiffer C, 2000). But sometimes important enzymes do not bring about a major change and an enzyme rarely occurring causes a major physiological change (Schmitz-Peiffer C, 2000), Hwang et al, 2007). Hence it is important to identify such enzymes and study their effects in metabolism.</p>
			</sec>
	</sec>
	<sec>
		<title>Current Strategies</title>
			<sec>
				<title>To Detect Gene Expression Level Changes</title>
					<p>Identifying genes and pathways associated with diseases such as cancer has been a subject of considerable research in recent years in the area of bioinformatics and computational biology (Speed et al, 2002; Herzyk et al, 2004). Many
statistical learning techniques such as support vector machines, the relevance vector machines (RVM), LASSO, and sparse logistic regression have been applied to this problem (Krogh et al, 1996). Such algorithms fulfill two common goals, viz., to distinguish cancer and non-cancer patients with the highest possible accuracy and to identify a small subset of genes that are highly differentiated in different classes and to associate gene expression patterns with disease status (Breitling et al, 2004). In gene selection, when genes share the same biological pathway, the correlation between them can be high and those genes form a group
(Breitling et al, 2004). The ideal gene selection methods eliminate the trivial genes and automatically include the whole group genes into the model once one gene among them is selected. Most importantly, almost all of the current methods are biased towards selecting those genes that display the most pronounced expression differences (Speed et al, 2002). Such methods select genes using purely statistical criteria (either rank score or classification accuracy) and this selection is thought to reflect their relative importance. Quite often, a certain number of genes with the smallest pvalues or highest prediction accuracy are finally selected,
while most biologists recognize that the magnitude of differential expression does not necessarily indicate biological significance (Breitling et al, 2004).</p>
<p>Although there is ongoing research to incorporate prior biological knowledge, such as partially known pathways in gene selection, to the best of our knowledge, there is no efficient method to hunt the upstream regulatory genes in gene selection and pathway discovery (Speed et al, 2002). Magnitude of differential expression does not necessarily indicate biological significance. Small changes can cause major physiological consequences if the protein encoded by the gene plays a catabolic role. Regulatory genes show low expression changes but they are highly correlated with the downstream highly expressed genes (Brown et al, 1997).
No efficient method to hunt upstream regulatory genes in gene selection exists, therefore an algorithm was developed which uses Lp norm regularization which is equivalent to super Laplace prior over the model parameters (Herzyk et al, 2004). Both the generalization ability of the model and the scarcity achieved are critically dependent on the value of a regularized parameter, which has to be carefully tuned to the best performance. This best parameter can only be found through cross-validation and computationally intensive search (Breitling et al, 2004). Various methods have been proposed for identifying the small subsets of genes through Lp regularized Bayesian logistic regression and then define a novel similarity measure to identify the regularized
genes that are highly correlated with each gene in the subset (Krogh et al, 1996).</p>
			</sec>
			<sec>
				<title>To Detect Nucleosome Positioning</title>
					<p>The positions of nucleosomes play important roles in diverse cellular processes that rely on access to genomic DNA, including DNA replication, recombination, repair, transcription, chromosome segregation, and cell division (Lorch
et al, 1999). Recent studies have used DNA sequence features to predict genome-wide nucleosome positions with modest success, confirming that nucleosome positioning is partially encoded in the genomic DNA sequence. It has become clear that nucleosome positions are highly dynamic. Previous computational methods have predicted static nucleosome positions using DNA sequences with nucleosome formation or inhibition signals. However, more information besides the intrinsic DNA sequence is required to model
nucleosome positioning dynamics. Till date, there has been no report on computational identification of dynamic nucleosome positions (Krogh et al, 1996). A novel computational approach was put forward for identifying dynamic nucleosome positioning at gene promoters on the base of dynamic transcriptional interaction and genomic sequence information. The predictions are in good agreement with experimentally determined nucleosome occupancy available in three cellular conditions. A recent study has used nucleosome occupancy information to assist identification of transcription factor (TF) binding sites (Lorch et al, 1999).</p>
<p>K-means clustering can also be used for detecting nucleosome positioning. Average nucleosome occupancy profile at promoters for each gene cluster can be found out and then pair-wise Euclidean distances can be calculated among these average profiles. The resulting distance reflects the degree of difference between the nucleosome occupancy profiles of two gene clusters (Caisheng et al, 2009). Based on dynamic transcriptional interaction and genomic sequence information, various computational approaches have been developed for identifying dynamic nucleosome positioning at promoters . Given gene expression data during one cellular
condition, DNA sequences at gene promoters, and known position weight matrixes (PWMs) that correspond to candidate TFs, could be used to identify nucleosome positioning dynamics during the condition.</p>
			</sec>
			<sec>
				<title>Graph Theory and Analysis for Detecting Enzyme Alterations</title>
					<p>A metabolic pathway may be represented computationally using a combination of statistical methods and programming methods (Speed et al, 2002). Data-sets are obtained from a central repository database which consists of metabolic
pathway information, from which graph-based representations are plotted (Vilaprinyo et al, 2006; Niyogi et al, 2005). These may also be used to represent the metabolic pathway and the altered enzymes could be identified through this method (Agarwal et al, 2006; Aleman-Meza et al, 2006; Vilaprinyo et al, 2006). Nodes may be represented are the compounds or the metabolites of the reaction or enzymes and edges may act as reaction links. Each node would contain a dataset which would help us segregate the list of similar entities. Statistical calculations such as, p-value calculation, matrix methods could be performed to rank the nodes (Krogh et al, 1996; Curto et al, 1997).</p>
<p>Ranking of the nodes help in arranging the datasets according to the enzymes causing maximum change in the pathway (Agarwal S, 2006; Kanehisa M, 2000; Vilaprinyo et al, 2006). The graph can be modeled on the basis of the calculated values which help identify the dataset (enzymes) causing maximum pathway changes. The nodes in the graph are ranked in order to indentify the connections and the link in the pathway. A method of ranking nodes comprises the steps of analyzing each node with respect to at least one
criterion, and assigning an intrinsic score to each node based on the analysis; identifying links between the nodes and generating a ranking score for each node based on the intrinsic scores of nodes linked therewith (Herzyk et al, 2004). Ranking of directed an undirected graphs would vary. According to the rank score the nodes are arranged according their particular rank. Weights are assigned to these ranked nodes by using matrix methods (Agarwal S, 2006; Kanehisa et al, 2000; Niyogi et al, 2005).</p>
			</sec>
			<sec>
				<title>A Bayesian method for identifying missing enzymes in predicted metabolic pathway databases</title>
					<p>The Pathologic program constructs Pathway/Genome databases by using a genome&rsquo;s annotation to predict the set of metabolic pathways present in an organism. Pathologic determines the set of reactions composing those pathways from the enzymes annotated in the organism&rsquo;s genome. Most annotation efforts fail to assign function to 40&ndash;60% of sequences. In addition, large numbers of sequences may have non-specific annotations (e.g., thiolase family protein). Pathway holes occur when a genome appears to lack the enzymes needed to catalyze reactions in a pathway [Pagano Marcello et al,1993]. If a protein has not been assigned a specific function during the annotation process, any reaction catalyzed by that protein will appear as a missing enzyme or pathway hole in a Pathway/Genome database. As a result of this a method has been developed that efficiently combines homology and pathway-basedevidence to identify candidates for filling pathway holes in Pathway/ Genome databases [Spang R etal 1998, Beutler. E, 1984]. The program not only identifies potential candidate sequences for pathway holes, but combines data from multiple, heterogeneous sources to assess the likelihood that a candidate has the required function. The algorithm emulates the manual sequence annotation process, considering not only evidence from homology searches, but also considering evidence from genomic context (i.e., is the gene part of an operon?) and functional context (e.g., are there functionally-related genes nearby in the genome?) to determine the posterior belief that a candidate has the required function. The method can be applied across an entire metabolic pathway network and is generally applicable to any pathway database[Pearl Judea,1988]. The program uses a set of sequences encoding the required activity in other genomes to identify candidate proteins in the genome of interest, and then evaluates each candidate by using a simple Bayesian classifier to determine the probability that the candidate has the desired 
function. 71% precision at a probability threshold of 0.9 during cross-validation was achieved using known reactions in computationally-predicted pathway databases. After applying this method to 513 pathway holes in 333 pathways from three Pathway/Genome databases, the number of complete pathways was increased by 42%. The pathway hole filler can be used not only to increase the utility of Pathway/ Genome databases to both experimental and computational researchers, but also to improve predictions of protein function.</p>
			</sec>
			<sec>
				<title>Influence of Metabolic Network Structure and Function on Enzyme Evolution</title>
				<p>Most studies of molecular evolution are focused on individual genes and proteins. However, understanding the design principles and evolutionary properties of molecular networks requires a system-wide perspective. In the present work we connect molecular evolution on the gene level with system properties of a cellular metabolic network. In contrast to protein interaction networks, where several previous studies investigated the molecular evolution of proteins, metabolic networks have a relatively well-defined global function [Saher G et al,2009]. The ability to consider fluxes in a metabolic network allows us to relate the functional role of each enzyme in a network to its rate of evolution. The result of an experiment based on the yeast metabolic network, demonstrate that important evolutionary processes, such as the fixation of single nucleotide mutations, gene
duplications, and gene deletions, are influenced by the structure and function of the network [Erb, R. S. et al,1999]. Specifically, central and highly connected enzymes evolve more slowly than less connected enzymes. Also, enzymes carrying high metabolic fluxes under natural biological conditions experience higher evolutionary constraints. Genes encoding enzymes with high connectivity and high metabolic flux have higher chances to retain duplicates in evolution [Watts D J et al,1998, Jeong H et al ,1999]. In contrast to protein interaction networks, highly connected enzymes are no more likely to be essential compared to less connected enzymes. The presented analysis of evolutionary constraints, gene duplication, and essentiality demonstrates that the structure and function of a metabolic network shapes the evolution of its enzymes. The results underscore the need for systems-based approaches in studies of molecular evolution.</p>
			</sec>
			<sec>
				<title>Quantitative, Scalable Discrete Event Simulation of Metabolic Pathways</title>
					<p>DMSS(Discrete Metabolic Simulation System) seeks to create a scalable, object oriented, quantitative discrete event simulation framework based on a model of metabolic pathways as a graph (or network) of reactions. The system will
facilitate the creation and evaluation of models through the simulation of experiments based on these models. Similar approaches have been used to model metabolic pathways, in particular the approach based on Petri nets (a type of graph) [Ehlde, M et al 1995]. This method has been suggested for modelling of pathways and for use in discrete{event simulation systems. Other approaches to
modelling and simulating metabolic pathways include dfferential equation and Metabolic Control [Farza, M et al ,1992] Theory systems, for example. In addition, there have been a number of attempts to use knowledge{based and reasoning systems, for both qualitative and quantitative simulation .DMSS is designed to be used by bench biologists, allowing complex biologicalmodels to be expressed in biological terms (for example, as metabolite concentrations or chemical reactions), rather than as sets of obscure equation constants [Erb, R. S,1999]. That is, the biological relevance of the input model, simulation process and thus the results is imperative. The first step to achieving this goal is the ability to create biologically relevant models, hence a model must
be defined in terms of biological constants [Pagano Marcello et al ,1993]. The system should not require this knowledge to be described in mathematical terms, particularly not with respect to time. In contrast, the constants used in differential equation based mathematical models arise as a consequence of requiring the formulae to describe the reaction model for a range of reaction related parameters. Additionally, the stability and sensitivity of the equations for any given parameters need to be determined and accounted for, as these are often artifacts in such equation{based simulation systems. DMSS seeks to simulate non trivial systems, an important issue for metabolic simulation systems ones which are considered problematic for other techniques.The simulation
of complex reaction systems can contribute to our understanding of biological systems and be an experimental platform for in silico biology.</p>
			</sec>
			<sec>
				<title>Ranking on Graph Data</title>
					<p>In ranking, one is given examples of order relationships among objects, and the goal is to learn from these examples a real-valued ranking function that induces a ranking or ordering over the object space [Agarwal, S , 2005]. We consider
the problem of learning such a ranking function when the data is represented as a graph, in which vertices correspond to objects and edges encode similarities between objects. Building on recent developments in regularization theory for graphs and corresponding Laplacian-based methods for classification, we develop an algorithmic framework for learning ranking functions on graph data [Belkin, M , 2004]. We provide generalization guarantees for our algorithms via recent results based on the notion of algorithmic stability, and give experimental evidence of the potential benefits of our framework.</p>
			</sec>
	</sec>
	<sec>
		<title>Application in Pathway Engineering</title>
			<p>Graph theory and other statistical measures can be used to detect the enzyme alterations using computational and
statistical methods. In order to detect the damages present in the pathway the graph theory plays a major role in representing the metabolic pathway and analyzing the connections between metabolites and their formation (Agarwal S, 2006; Niyogi et al, 2005; Niyogi et al, 2004). This can be done from the pre-collected data repository. For detection of enzymatic changes the nodes and edges are connected and the degree of the enzymes and metabolites is calculated, which would further help find the rank of the same. From the ranks, one could find out occurrence of the enzymes and their positions in the pathway by comparing the ranks with the respective nodes. This study would help to categorize the enzymes that bring about maximum physiological change and hence detect the altered enzymes in a metabolic pathway. The detection of altered enzymes in a metabolic pathway may help find cure for many metabolic disorders.</p>
<p>Ranking may also help to analyze the differential gene expression from the microarray data available (Speed et al, 2002; Herzyk et al, 2004). Graph is plotted and the ranking is done based on highly differentially expressed genes. The ranking of a graph also depends on the degree of the graph. The degree of the graph is basically the number of subnodes a particular node is connected to (Herzyk et al, 2004; Breitling et al, 2004). The number of nodes having the same degree is cumulatively placed under the same rank.</p>
	</sec>
	<sec>
		<title>Conclusion &amp; Future Perspective</title>
			<p>Damages in metabolic pathways can be identified by representing the metabolic pathway in the form of graphs. A
graph based analysis which implements statistical methods help represent metabolic pathways and understand the function of a particular enzyme and the changes caused in the pathway due to the absence of the enzyme. Integration of statistical and computational methods has helped formulate an algorithm to determine the changes caused in a metabolic pathway. The detection of these damages can help understand the disorders caused due to these damages.</p>
	</sec>
	</body>
	<back>
	<ref-list>
	   <title>References</title>
	         <ref id="r1">
			 <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Abedin</surname>
						  <given-names>MJ</given-names>
					</name>
					<name>
					      <surname>Wang</surname>
						  <given-names>D</given-names>
					</name>
					<name>
					      <surname>McDonnell</surname>
						  <given-names>MA</given-names>
					</name>
					<name>
					      <surname>Lehmann</surname>
						  <given-names>U</given-names>
					</name>
					<name>
					      <surname>Kelekar</surname>
						  <given-names>A</given-names>
					</name>
					</person-group>
					<year>2007</year>
					<article-title>Autophagy delays apoptotic death in breast cancer cells following DNA damage</article-title>
					<source>Cell Death Differ</source>
					<volume>14</volume>					
					<fpage>500</fpage>
					<lpage>510</lpage>
			</citation>
			</ref>
			<ref id="r2">
			 <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Agarwal</surname>
						  <given-names>S</given-names>
					</name>
					</person-group>
					<year>2006</year>
					<article-title>Ranking on graph data</article-title>
					<source>ICML</source>				
					<fpage>pp25</fpage>
					<lpage>pp32</lpage>
			</citation>
			</ref>
			<ref id="r3">
			 <citation citation-type="confproc">
			         <person-group>
					 <name>
					      <surname>Agarwal</surname>
						  <given-names>A</given-names>
					</name>
					<name>
					      <surname>Chakrabarti</surname>
						  <given-names>S</given-names>
					</name>
					<name>
					      <surname>Aggarwal</surname>
						  <given-names>S</given-names>
					</name>
					</person-group>
					<year>2006</year>
					<article-title>Learning to rank networked entities</article-title>
					<conf-name>SIGKDD Conference</conf-name>
					<fpage>pp14</fpage>
					<lpage>pp23</lpage>
			</citation>
			</ref>
			<ref id="r4">
			 <citation citation-type="confproc">
			         <person-group>
					 <name>
					      <surname>Agarwal</surname>
						  <given-names>S</given-names>
					</name>
					<name>
					      <surname>Niyogi</surname>
						  <given-names>P</given-names>
					</name>			
					</person-group>
					<year>2005</year>
					<article-title>Stability and generalization of bipartite ranking algorithms</article-title>								
					<conf-name>Proceedings of the 18th Annual Conference on Learning Theory</conf-name>				
			</citation>
			</ref>
			<ref id="r5">
			  <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Aita</surname>
						  <given-names>VM</given-names>
					</name>
					<name>
					      <surname>Liang</surname>
						  <given-names>XH</given-names>
					</name>		
					<name>
					      <surname>Murty</surname>
						  <given-names>VV</given-names>
					</name>		
					<name>
					      <surname>Pincus</surname>
						  <given-names>DL</given-names>
					</name>	
					<name>
					      <surname>Yu</surname>
						  <given-names>W</given-names>
					</name>	<etal/>		
					</person-group>
					<year>1999</year>
					<article-title>Cloning and genomic organization of beclin 1, a candidate tumor suppressor gene on chromosome 17q21</article-title>
					<source>Genomics</source>
					<volume>59</volume>					
					<fpage>59</fpage>
					<lpage>65</lpage>
			</citation>
			</ref>
			<ref id="r6">
			 <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Albertson</surname>
						  <given-names>DG</given-names>
					</name>			
					</person-group>
					<year>2006</year>
					<article-title>Gene amplification in cancer</article-title>
					<source>Trends Genet</source>
					<volume>22</volume>					
					<fpage>447</fpage>
					<lpage>455</lpage>
			</citation>
			</ref>
			<ref id="r7">
			 <citation citation-type="confproc">
			         <person-group>
					 <name>
					      <surname>Aleman-Meza</surname>
						  <given-names>B</given-names>
					</name>		
					</person-group>
					<year>2006</year>
					<article-title>Searching and ranking documents based on semantic relationships</article-title>								
					<conf-name>22nd International Conference on Data Engineering Workshops (ICDEW&rsquo;06)</conf-name>				
					<fpage>pp5</fpage>
			</citation>
			</ref>
			<ref id="r8">
			  <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Bartkova</surname>
						  <given-names>J</given-names>
					</name>
					<name>
					      <surname>Horejs&iacute;</surname>
						  <given-names>Z</given-names>
					</name>		
					<name>
					      <surname>Koed</surname>
						  <given-names>K</given-names>
					</name>		
					<name>
					      <surname>Kr&auml;mer</surname>
						  <given-names>A</given-names>
					</name>		
					<name>
					      <surname>Tort</surname>
						  <given-names>F</given-names>
					</name>		<etal/>
					</person-group>
					<year>2005</year>
					<article-title>DNA damage response as a candidate anti-cancer barrier in early human tumorigenesis</article-title>
					<source>Nature</source>
					<volume>434</volume>					
					<fpage>864</fpage>
					<lpage>870</lpage>
			</citation>
			</ref>
			<ref id="r9">
			  <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Beljanski</surname>
						  <given-names>V</given-names>
					</name>
					<name>
					      <surname>Marzilli</surname>
						  <given-names>LG</given-names>
					</name>		
					<name>
					      <surname>Doetsch</surname>
						  <given-names>PW</given-names>
					</name>		
					</person-group>
					<year>2004</year>
					<article-title>DNA Damage- Processing Pathways Involved in the Eukaryotic Cellular Response to Anticancer DNA Cross-Linking Drugs</article-title>
					<source>Mol Pharmacol</source>
					<volume>65</volume>					
					<fpage>1496</fpage>
					<lpage>1506</lpage>
			</citation>
			</ref>
			<ref id="r10">
			 <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Belkin</surname>
						  <given-names>M</given-names>
					</name>
					 <name>
					      <surname>Matveeva</surname>
						  <given-names>I</given-names>
					</name>
					<name>
					      <surname>Niyogi</surname>
						  <given-names>P</given-names>
					</name>					
					</person-group>
					<year>2004</year>
					<article-title>Regularization and semi-supervised learning on large graphs</article-title>								
					<conf-name>Proceedings of the 17th Annual Conference on Learning Theory</conf-name>			
			</citation>
			</ref>
			<ref id="r11">
			  <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Breitling</surname>
						  <given-names>R</given-names>
					</name>
					<name>
					      <surname>Amtmann</surname>
						  <given-names>A</given-names>
					</name>		
					<name>
					      <surname>Herzyk</surname>
						  <given-names>P</given-names>
					</name>			
					</person-group>
					<year>2004</year>
					<article-title>Graph-based Iterative Group Analysis (iGA): A simple tool to enhance sensitivity and facilitate interpretation of microarray experiments</article-title>
					<source>BMC Bioinformatics</source>
					<volume>5</volume>					
					<fpage>100</fpage>
			</citation>
			</ref>
			<ref id="r12">
			  <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Breitling</surname>
						  <given-names>R</given-names>
					</name>
					<name>
					      <surname>Armengaud</surname>
						  <given-names>P</given-names>
					</name>		
					<name>
					      <surname>Amtmann</surname>
						  <given-names>A</given-names>
					</name>		
					<name>
					      <surname>Herzyk</surname>
						  <given-names>P</given-names>
					</name>			
					</person-group>
					<year>2004</year>
					<article-title>Rank products: A simple, yet powerful, new method to detect differentially regulated genes in replicated microarray experiments</article-title>
					<source>FEBS Lett</source>
					<volume>573</volume>					
					<fpage>83</fpage>
					<lpage>92</lpage>
			</citation>
			</ref>
			<ref id="r13">
			 <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Cordwell</surname>
						  <given-names>SJ</given-names>
					</name>										
					</person-group>
					<year>1999</year>
					<article-title>Microbial genomes and &ldquo;missing&rdquo; enzymes: redefining biochemical pathways</article-title>										
					<source>Archives Microbiology</source> 
					<volume>172</volume>
					<fpage>269</fpage>
					<lpage>279</lpage>					
			</citation>
			</ref>
			<ref id="r14">
			 <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Curto</surname>
						  <given-names>R</given-names>
					</name>
					<name>
					      <surname>Voit</surname>
						  <given-names>EO</given-names>
					</name>	
					<name>
					      <surname>Sorribas</surname>
						  <given-names>A</given-names>
					</name>
					<name>
					      <surname>Cascante</surname>
						  <given-names>M</given-names>
					</name>										
					</person-group>
					<year>1997</year>
					<article-title>Validation and steady-state analysis of a power-law model of purine metabolism in man</article-title>										
					<source>Biochem J</source>
					<volume>324</volume>
					<fpage>761</fpage>
					<lpage>775</lpage>					
			</citation>
			</ref>
			<ref id="r15">
			 <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Dai</surname>
						  <given-names>Z</given-names>
					</name>
					<name>
					      <surname>Dai</surname>
						  <given-names>X</given-names>
					</name>		
					<name>
					      <surname>Xiang</surname>
						  <given-names>Q</given-names>
					</name>		
					<name>
					      <surname>Feng</surname>
						  <given-names>J</given-names>
					</name>	
					<name>
					      <surname>Deng</surname>
						  <given-names>Y</given-names>
					</name>		<etal/>	
					</person-group>
					<year>2009</year>
					<article-title>Transcriptional interaction-assisted identification of dynamic nucleosome positioning</article-title>
					<source>BMC Bioinformatics</source>
					<volume>1</volume>					
					<fpage>S31</fpage>
			</citation>
			</ref>
			<ref id="r16">
			  <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>DeRisi</surname>
						  <given-names>JL</given-names>
					</name>
					<name>
					      <surname>Iyer</surname>
						  <given-names>VR</given-names>
					</name>		
					<name>
					      <surname>Brown</surname>
						  <given-names>PO</given-names>
					</name>		
					</person-group>
					<year>1997</year>
					<article-title>Exploring the metabolic and genetic control of gene expression on a genomic scale</article-title>
					<source>Science</source>
					<volume>278</volume>					
					<fpage>680</fpage>
					<lpage>686</lpage>
			</citation>
			</ref>
			<ref id="r17">
			  <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Guo</surname>
						  <given-names>Z</given-names>
					</name>	
					</person-group>
					<year>2008</year>
					<article-title>Intramyocellular lipid kinetics and insulin resistance</article-title>
					<source>Med. Hypotheses</source>
					<volume>70</volume>					
					<fpage>625</fpage>
					<lpage>629</lpage>
			</citation>
			</ref>
			<ref id="r18">
			  <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Herwig</surname>
						  <given-names>R</given-names>
					</name>
					<name>
					      <surname>Aanstad</surname>
						  <given-names>P</given-names>
					</name>		
					<name>
					      <surname>Clark</surname>
						  <given-names>M</given-names>
					</name>		
					<name>
					      <surname>Lehrach</surname>
						  <given-names>H</given-names>
					</name>			
					</person-group>
					<year>2001</year>
					<article-title>Statistical evaluation of differential expression on cDNA nylon arrays with replicated experiments</article-title>
					<source>Nucleic Acids Res</source>
					<volume>29</volume>					
					<fpage>E117</fpage>
			</citation>
			</ref>
			<ref id="r19">
			  <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Huang</surname>
						  <given-names>W</given-names>
					</name>
					<name>
					      <surname>Wang</surname>
						  <given-names>P</given-names>
					</name>		
					<name>
					      <surname>Liu</surname>
						  <given-names>Z</given-names>
					</name>		
					<name>
					      <surname>Zhang</surname>
						  <given-names>L</given-names>
					</name>			
					</person-group>
					<year>2009</year>
					<article-title>Identifying disease associations via genome-wide association studies</article-title>
					<source>BMC Bioinformatics</source>
					<volume>1</volume>					
					<fpage>S68</fpage>
			</citation>
			</ref>
			<ref id="r20">
			 <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Kohavi</surname>
						  <given-names>R</given-names>
					</name>
					<name>
					      <surname>John</surname>
						  <given-names>GH</given-names>
					</name>										
					</person-group>
					<year>1997</year>
					<article-title>Wrappers for feature subset selection</article-title>										
					<source>Artificial Intelligence</source>
					<volume>97</volume>
					<fpage>273</fpage>
					<lpage>324</lpage>					
			</citation>
			</ref>
			<ref id="r21">
			 <citation citation-type="confproc">
			         <person-group>
					 <name>
					      <surname>Latourrette</surname>
						  <given-names>M</given-names>
					</name>										
					</person-group>
					<year>2000</year>
					<article-title>Toward an explanatory similarity measure for nearest-neighbor classification</article-title>
					<conf-name>In: ECML '00: Proceedings of the 11th European Conference on Machine Learning</conf-name>
					<conf-loc>London UK</conf-loc>					
					<fpage>238</fpage>
					<lpage>245</lpage>										
			</citation>
			</ref>
			<ref id="r22">
			 <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Liu</surname>
						  <given-names>B</given-names>
					</name>
					<name>
					      <surname>Abbass</surname>
						  <given-names>HA</given-names>
					</name>
					<name>
					      <surname>McKay</surname>
						  <given-names>B</given-names>
					</name>										
					</person-group>
					<year>2004</year>
					<article-title>Classification Rule Discovery with Ant Colony Optimization</article-title>										
					<source>IEEE Computational Intelligence Bulletin</source>
					<volume>3</volume>
					<fpage>31</fpage>
					<lpage>35</lpage>					
			</citation>
			</ref>
			<ref id="r23">
			 <citation citation-type="book">
			         <person-group>
					 <name>
					      <surname>Liu</surname>
						  <given-names>H</given-names>
					</name>
					<name>
					      <surname>Motoda</surname>
						  <given-names>H</given-names>
					</name>															
					</person-group>
					<year>1998</year>
					<article-title>Feature Selection for Knowledge Discovery and Data Mining</article-title>										
					<publisher-name>Boston: Kluwer Academic Publishers</publisher-name>										
			</citation>
			</ref>
			<ref id="r24">
			 <citation citation-type="journal">
			         <person-group>
					 <name>
					      <surname>Maniezzo</surname>
						  <given-names>V</given-names>
					</name>
					<name>
					      <surname>Colorni</surname>
						  <given-names>A</given-names>
					</name>															
					</person-group>
					<year>1999</year>
					<article-title>The Ant System Applied to the Quadratic Assignment Problem</article-title>										
					<source>IEEE Transaction on Knowledge and Data Engineering</source>
					<volume>11</volume>
					<fpage>pp769</fpage>
					<lpage>778</lpage>					
			</citation>
			</ref>
			<ref id="r25">
			 <citation citation-type="book">
			         <person-group>
					 <name>
					      <surname>Mitchell</surname>
						  <given-names>T</given-names>
					</name>																				
					</person-group>
					<year>1996</year>
					<article-title>Machine Learning</article-title>										
					<publisher-name>McCraw Hill</publisher-name>										
			</citation>
			</ref>			
	</ref-list>
	</back>			 					  						 
</article>