Back

Mapping the phenotypic landscape of a transcriptional repressor using Deep Mutational Scanning and Growth-based Quantitative Sequencing

Jansen, Z.; Le, X.; Wei, Q.; Kulhanek, D. L.; Alperovich, N.; Vasilyeva, O. B.; Gilmour, A. R.; Ross, D.; Thyer, R.

2025-12-12 synthetic biology
10.64898/2025.12.11.693801 bioRxiv
Show abstract

CymR is a TetR-family transcriptional repressor that recognizes a well-defined operator sequence in the promoter PcymRC. The native ligand cumate and several structurally related aromatic acids bind at an allosteric site and induce a conformational change in CymR, resulting in release from the DNA operator and de-repression of the promoter. The amino acid residues that contribute to these core functions have not been mapped, nor has the protein been subjected to extensive mutagenesis to modify its function. Here, for the first time, we integrate Deep Mutational Scanning (DMS) with Growth-based Quantitative Sequencing (GROQ-Seq) to evaluate a comprehensive phenotypic landscape of CymR variants, including single amino acid insertions and deletions. We measure this library across a concentration gradient of small molecule inducers to construct an induction curve for all library members. From this analysis, we identify amino acids throughout the protein that are essential for repressor function and discover several mutations that improve the sensitivity of CymR to the ligand perillic acid. In addition, rarely investigated insertion mutants are revealed to be a key driver of novel phenotypes, including several regions of CymR where insertions result in an inverted phenotype and the isolation of variants exhibiting an unusual band-stop phenotype. Graphical Abstract O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=81 SRC="FIGDIR/small/693801v1_ufig1.gif" ALT="Figure 1"> View larger version (30K): org.highwire.dtl.DTLVardef@121b16org.highwire.dtl.DTLVardef@b0478aorg.highwire.dtl.DTLVardef@128f2ecorg.highwire.dtl.DTLVardef@164a426_HPS_FORMAT_FIGEXP M_FIG C_FIG

Published in Nucleic Acids Research (predicted rank #3) · training set

Matching journals

The top 2 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.