SearcharxivSearch

arXiv subjects

Tom Röschinger

Publications and source records attributed to Tom Röschinger.

3 recordsLinked to original sources

Informational blueprints reveal condition-dependent gene regulatory architectures

While coding regions in the genome have a direct interpretation in terms of protein products, significant fractions are non-coding and yet control essential biological functions. Unlike the genetic code, there is no "lookup table" that identifies where regulatory proteins, known as transcription factors (TFs), bind. Here, we extract these binding sites by distilling sequences of nucleotide letters into collective coordinates (hyperletters) representing the binding sites that are active under specific environmental conditions. Going beyond local information footprints between individual bases and expression levels, our $\textit{information blueprint}$ algorithm compresses the global information by optimising filters that simultaneously scan an entire promoter sequence. Inspired by renormalisation-group techniques, we identify TF binding sites as coarse-grained variables combining groups of correlated mutations with the highest collective impact on gene expression. We validate our approach on experimental data for $\textit{E. coli}$ and discover novel regulatory elements illustrating its deployment at scale across growth conditions.

q-bio.GN

The Environment-Dependent Regulatory Landscape of the E. coli Genome

All cells respond to changes in both their internal milieu and the environment around them through the regulation of their genes. Despite decades of effort, there remain huge gaps in our knowledge of both the function of many genes (the so-called y-ome) and how they adapt to changing environments via regulation. Here we describe a joint experimental and theoretical dissection of the regulation of a broad array of over 100 biologically interesting genes in E. coli across 39 diverse environments, permitting us to discover the binding sites and transcription factors that mediate regulatory control. Using a combination of mutagenesis, massively parallel reporter assays, mass spectrometry and tools from information theory and statistical physics, we go from complete ignorance of a promoter's environment-dependent regulatory architecture to predictive models of its behavior. As a proof of principle of the biological insights to be gained from such a study, we chose a combination of genes from the y-ome, toxin-antitoxin pairs, and genes hypothesized to be part of regulatory modules; in all cases, we discovered a host of new insights into their underlying regulatory landscape and resulting biological function.

q-bio.GN

Adaptive ratchets and the evolution of molecular complexity

Biological systems have evolved to amazingly complex states, yet we do not understand in general how evolution operates to generate increasing genetic and functional complexity. Molecular recognition sites are short genome segments or peptides binding a cognate recognition target of sufficient sequence similarity. Such sites are simple, ubiquitous modules of sequence information, cellular function, and evolution. Here we show that recognition sites, if coupled to a time-dependent target, can rapidly evolve to complex states with larger code length and smaller coding density than sites recognising a static target. The underlying fitness model contains selection for recognition, which depends on the sequence similarity between site and target, and a uniform cost per unit of code length. Site sequences are shown to evolve in a specific adaptive ratchet, which produces selection of different strength for code extensions and compressions. Ratchet evolution increases the adaptive width of evolved sites, accelerating the adaptation to moving targets and facilitating refinement and innovation of recognition functions. We apply these results to the recognition of fast-evolving antigens by the human immune system. Our analysis shows how molecular complexity can evolve as a collateral to selection for function in a dynamic environment.

q-bio.PE