SearcharxivSearch

arXiv · 2012.06222

Glycolytic pyruvate kinase moonlighting activities in DNA replication initiation and elongation

Abstract

Cells have evolved a metabolic control of DNA replication to respond to a wide range of nutritional conditions. Accumulating data suggest that this poorly understood control depends, at least in part, on Central Carbon Metabolism (CCM). In Bacillus subtilis , the glycolytic pyruvate kinase (PykA) is intricately linked to replication. This 585 amino-acid-long enzyme comprises a catalytic (Cat) domain that binds to phosphoenolpyruvate (PEP) and ADP to produce pyruvate and ATP, and a C-terminal domain of unknown function. Interestingly, the C-terminal domain termed PEPut interacts with Cat and is homologous a domain that, in other metabolic enzymes, are phosphorylated at a conserved TSH motif at the expense of PEP and ATP to drive sugar import and catalytic or regulatory activities. To gain insights into the role of PykA in replication, DNA synthesis was analyzed in various Cat and PEPut mutants grown in a medium where the metabolic activity of PykA is dispensable for growth. Measurements of replication parameters ( ori/ter ratio, C period and fork speed) and of the pyruvate kinase activity showed that PykA mutants exhibit replication defects resulting from side chain modifications in the PykA protein rather than from a reduction of its metabolic activity. Interestingly, Cat and PEPut have distinct commitments in replication: while Cat impacts positively and negatively replication fork speed, PEPut stimulates initiation through a process depending on Cat-PEPut interaction and growth conditions. Residues binding to PEP and ADP in Cat, stabilizing the Cat-PEPut interaction and belonging to the TSH motif of PEPut were found important for the commitment of PykA in replication. In vitro , PykA affects the activities of replication enzymes (the polymerase DnaE, helicase DnaC and primase DnaG) essential for initiation and elongation and genetically linked to pykA . Our results thus connect replication initiation and elongation to CCM metabolites (PEP, ATP and ADP), critical Cat and PEPut residues and to multiple links between PykA and the replication enzymes DnaE, DnaC and DnaG. We propose that PykA is endowed with a moonlighting activity that senses the concentration of signaling metabolites and interacts with replication enzymes to convey information on the cellular metabolic state to the replication machinery and adjust replication initiation and elongation to metabolism. This defines a new type of replication regulator proposed to be part of the metabolic control that gates replication in the cell cycle.

Explore related subjects

Keep this discovery

BibTeXRIS

Steff Horemans, Matthaios Pitoulias, Alexandria Holland, Panos Soultanas, Laurent Janniere. 2020-12-11. Glycolytic pyruvate kinase moonlighting activities in DNA replication initiation and elongation. https://arxiv.org/abs/2012.06222

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Thermodynamic and Statistical Signatures of Modality Changes in Concentration Distributions Driven by Stochastic Switching Between Two Activity States

Stochastic switching between gene expression states, coupled with production and degradation dynamics, governs the accumulation of mRNA and proteins in cells. The concentrations of these accumulated entities dictate the phenotypic distribution of genetically identical cells. The underlying accumulation dynamics are well-captured by a two-state promoter switching model, with statistical and thermodynamic properties quantified via the Fano factor and entropy production rates. However, how these measures correlate with concentration distributions and their shifts under varying kinetic parameters remains largely unexplored. To this end, we use chemical master equations to study a generalized model of mRNA accumulation dynamics in the presence of stochastic switching between two activity states and state-dependent production and degradation rates. We derive exact expressions for the steady-state probability distribution and analytically compute the mean concentration, Fano factor, and entropy production rate (EPR). Simplifying these expressions, we identify contributions arising from stochastic switching rates and relaxation dynamics toward equilibrium in each activity state. Next, using our theoretical results, we characterize the variation in the Fano factor and EPR as a function of mean expression during modality changes of the distributions mediated by the variation of switching rates. We also identify the conditions in kinetic parameters that achieve the highest Fano factor and entropy production rates. Our findings establish a generalized framework for examining stochastic accumulation dynamics, clarifying how kinetic parameters dictate molecular distributions, noise, and dissipation. These insights extend readily to broader contexts coupling stochastic switching with accumulation, including protein burst dynamics, phenotype-switching-mediated drug intake, and queuing theory.

q-bio.MN

Are You Learning Biological Signal or Shortcuts? Auditing and Mitigating Bias in Protein-Protein Interaction Datasets

Protein-protein interaction (PPI) databases do not faithfully reflect biological realities. Instead, they are influenced by study and technical biases that distort certain protein and interaction attributes. Machine learning models can exploit these as learning shortcuts if the negative dataset is not constructed with care. So far, the shortcuts introduced during PPI dataset construction have only been examined in isolation. Here, we systematically characterize both reported and, to our knowledge, previously unreported biases in PPI datasets that lead machine learning models to learn shortcuts instead of biological signal. We analyze HIPPIE, IntAct, and STRING, dedicated PPI databases, as well as two datasets derived from 3D-structural information in the Protein Data Bank (PDB). We show that random data splitting introduces strong topological shortcuts. When train-test protein overlap is removed, the resulting datasets still retain usable shortcuts stemming from self-interactions, taxonomic identity, and functional relatedness, whose prevalence interestingly depends on the data source. We further show that sampling negatives from a set of high-confidence non-interactors, an intuitively appealing choice, can amplify the shortcut stemming from functional relatedness. To detect and mitigate these biases, we provide an open Nextflow pipeline that combines similarity-aware, data-loss-minimizing dataset splitting with bias-minimizing negative sampling, both formulated as integer linear programs. Its key concept of quantifying biases to minimize them through optimization-based negative sampling can, in principle, be extended to any machine learning problem where the pool of negative candidates is much larger than the positives and is thus of interest also beyond PPI prediction.

q-bio.MN

Orchestra: Corroboration-Based Regulatory Candidate Discovery via Composed Bioinformatics MCP Agents

Orchestra composes two independently built bioinformatics MCP servers -- RegNetAgents, which infers gene regulatory network topology from ARACNe networks, and CASCADE, which supplies four independent evidence sources (LINCS knockdown, DepMap essentiality, super-enhancer status, DoRothEA transcription-factor confidence) -- into one multi-agent workflow exposed via the Model Context Protocol. Its central architectural claim is that requiring RegNetAgents' topology evidence and CASCADE's experimental evidence to agree on a candidate regulator yields a more trustworthy candidate than either alone -- not previously tested directly, since RegNetAgents' own validation asked only whether its candidate lists beat chance. We test this on the TCGA tumor-acquired regulator tier (regulators in a gene's tumor ARACNe network but absent from the GREmLN population-averaged baseline), selecting candidates by ARACNe mutual-information (MI) edge weight. On RegNetAgents' published BRCA/COAD focal-gene panel plus matched negative controls, agreement among at least 2 of the 4 CASCADE sources predicts OncoKB cancer-gene status among focal genes (odds ratio 2.89, Benjamini-Hochberg-adjusted p=0.0166) but not among negative controls (p=0.0721); a single source is not diagnostic for either group. The pattern replicates and strengthens in a third cancer type, STAD, on a separately constructed panel (odds ratio 5.82), and against an independently curated ground truth (the Sanger COSMIC Cancer Gene Census). MI edge weight is the strongest single predictor overall (p=0.0003); a logistic-regression likelihood-ratio test confirms corroboration adds value beyond it in both panels (p=0.0234; p=0.0001). Every experiment invokes Orchestra's real agentic entry point.

q-bio.MN