Language Selection

Get healthy now with MedBeds!
Click here to book your session

Protect your whole family with Orgo-Life® Quantum MedBed Energy Technology® devices.

Advertising by Adpathway

         

 Advertising by Adpathway

GHT-SELEX reveals strong sequence specificity in human transcription factors

1 hour ago 5

PROTECT YOUR DNA WITH QUANTUM TECHNOLOGY

Orgo-Life the new way to the future

  Advertising by Adpathway

For decades, molecular biologists have treated the DNA-binding preferences of transcription factors as one of the more settled chapters of gene regulation. Every textbook diagram shows these proteins docking onto short, well-defined DNA motifs, switching genes on or off with clean specificity. A new study published in Nature Methods upends that tidy picture. Using an upgraded experimental technique called GHT-SELEX, a research team led by Arttu Jolma, together with Adrian Hernandez-Corchado and A.W.H. Yang and colleagues, has found that many human transcription factors possess far higher intrinsic sequence specificity — and far more complex DNA-binding behavior — than anyone had previously measured. The work, published in the journal’s September 2026 issue as an article spanning pages 1775 to 1785, is already generating discussion among researchers who map gene regulatory networks, because it suggests that substantial portions of the published literature on transcription factor binding sites may be incomplete or even misleading.

Transcription factors are the master switches of the genome. In humans, roughly 1,600 of these proteins read the DNA sequence and decide, in concert with one another, which of our roughly 20,000 genes are active in any given cell. They do this by recognizing specific short stretches of DNA — typically 6 to 12 base pairs — known as binding motifs. Knowing exactly which sequence each factor prefers is foundational to nearly everything in genomics: predicting how genetic variants contribute to disease, engineering synthetic gene circuits, interpreting genome-wide association studies, and understanding how mutations in regulatory regions drive cancer. Yet despite the importance of these measurements, the standard methods used to derive them have long been recognized as imperfect.

The dominant approaches — including protein binding microarrays, conventional SELEX (Systematic Evolution of Ligands by Exponential Enrichment), and various high-throughput SELEX variants — share a common limitation: they typically measure binding under a single, fixed set of conditions, and they often rely on initial libraries whose sequence diversity constrains what can be discovered. A transcription factor that binds weakly, binds cooperatively, or requires specific flanking context can easily be mischaracterized. GHT-SELEX, the method introduced and deployed in the new study, was designed to close these gaps. The “GHT” in its name reflects an expanded high-throughput design that dramatically increases both the depth of sequencing and the diversity of the DNA library interrogated in each round of selection, allowing the technique to capture binding events that earlier methods would have missed entirely.

In a GHT-SELEX experiment, a vast library of random DNA sequences is incubated with a purified transcription factor. The sequences that the protein binds are separated from those it ignores — typically using techniques that pull down the protein-DNA complexes — and the bound sequences are then amplified and subjected to another round of selection. After several iterative cycles, the pool becomes enriched for high-affinity binding sequences, and deep sequencing reveals what the protein “chose.” The statistical analysis of these enriched sequences, combined with careful modeling of binding energies, allows researchers to reconstruct a precise portrait of the protein’s sequence preferences, including subtle dependencies between positions that simpler models cannot capture. The key innovation in this study lies in scaling this approach up and in applying it systematically to a large panel of human transcription factors under conditions designed to reveal their full behavioral repertoire.

The results were striking. A substantial number of the transcription factors examined displayed sequence specificity that is “unexpectedly high” — meaning their discrimination between favored and disfavored DNA sequences is far sharper than earlier assays had indicated. Where previous studies might have characterized a factor as recognizing a loose, degenerate motif, GHT-SELEX revealed that the protein in fact distinguishes finely between closely related sequences, tolerating only a narrow band of variation. This matters enormously for interpretation of genomic data. If a factor is actually highly specific but has been modeled as promiscuous, then computational predictions of where it binds across the genome — and which genetic variants might disrupt those bindings — will be systematically wrong.

Just as consequential is the second headline finding: complex DNA binding. Many of the factors did not behave as simple, independent-position recognizers at all. Instead, their binding depended on interactions between positions in the motif, on the spacing and orientation of multiple recognition elements, and in some cases on the ability to engage more than one DNA site at a time. Some factors showed evidence of dimeric binding on concatenated sites; others exhibited context-dependent preferences in which the sequence flanking the core motif altered what the core itself could be. These are exactly the kinds of behaviors that conventional motif models — which assume each position in a binding site contributes independently to binding affinity — cannot represent. The study’s data indicate that such “independent position” assumptions, baked into the position weight matrix models used ubiquitously in bioinformatics, fail for a meaningful fraction of human transcription factors.

The implications ripple outward across several fields. In regulatory genomics, motif scanning underlies algorithms such as those used to annotate transcription factor binding sites in the human genome and to interpret ENCODE-style functional element catalogs. If the underlying specificity models are too coarse, then hundreds of thousands of predicted binding sites may be false positives, while genuinely important sites — those that depend on complex, cooperative or context-sensitive binding — may be absent from the catalogs entirely. In medical genetics, fine-mapping studies that try to pinpoint which regulatory variant explains a disease association rely on motif disruption scores; sharper, more accurate specificity models should translate directly into better variant interpretation. In synthetic biology, engineers who design genetic circuits using natural transcription factors will now have better ground truth about what sequences their parts actually respond to.

The methodological advance is itself noteworthy. By combining an enlarged library design with high-depth sequencing and quantitative modeling, GHT-SELEX achieves a dynamic range that allows both very strong and comparatively weak binding interactions to be measured within the same experiment. This is important because biological reality rarely fits a single affinity: transcription factors in living cells encounter a spectrum of sites, and their functional occupancy depends on affinity, competition with other factors, chromatin context and concentration. Having an in vitro assay that can resolve this spectrum — rather than collapsing it into a single consensus motif — brings the measurement considerably closer to the biology.

The study also underscores how much of the human transcription factor repertoire remains incompletely characterized. Even for well-studied families — the homeodomains, bZIPs, bHLHs, nuclear receptors and zinc finger proteins that appear throughout the gene regulation literature — the new measurements revealed surprises, including preferences and binding modes not captured in existing databases such as JASPAR or HOCOMOCO. For factors previously annotated only by similarity to relatives, the new data provide direct experimental characterization, some of it quite different from what homology-based transfer would have predicted. The authors’ systematic approach, applying the same protocol across a broad panel of proteins, also makes the resulting dataset unusually consistent and comparable — a valuable resource for anyone building predictive models of gene regulation.

For the broader field, the paper is likely to prompt a re-evaluation of how transcription factor specificity is measured and modeled. Position weight matrices will not disappear overnight — they remain useful first approximations — but the study strengthens the case for higher-order models, such as dinucleotide models and deep-learning-based approaches, that can capture inter-positional dependencies. It also makes a strong argument for experimental rigor: the intrinsic preferences of a transcription factor, measured cleanly in vitro with sufficient dynamic range, can differ enough from legacy measurements to change biological conclusions. As researchers begin incorporating the new specificity data into genome-wide analyses, one thing seems certain: the grammar of gene regulation, long treated as a simple code of short motifs, is proving to be considerably richer — and considerably more precise — than the textbooks suggested.

Subject of Research: Human transcription factors and their intrinsic DNA sequence specificity and complex DNA-binding behavior, measured using a high-throughput in vitro selection method (GHT-SELEX)

Subject of Research: Biology

Article Title: GHT-SELEX demonstrates unexpectedly high intrinsic sequence specificity and complex DNA binding of many human transcription factors

Article References: Jolma, A., Hernandez-Corchado, A., Yang, A. W. H., Fathi, A., Laverty, K. U., Brechalov, A., Razavi, R., Albu, M., Zheng, H., The Codebook Consortium, Bucher, P., Deplancke, B., Fornes, O., Jan Grau, Grosse, I., Kolpakov, F. A., Makeev, V. J., Barazandeh, M., Deng, Z., … Hughes, T. R. (2026). GHT-SELEX demonstrates unexpectedly high intrinsic sequence specificity and complex DNA binding of many human transcription factors. Nature Methods, 23(9), 1775-1785. https://doi.org/10.1038/s41592-026-03177-9

Image Credits: AI Generated

DOI: 10.1038/s41592-026-03177-9

Keywords: GHT-SELEX, transcription factors, DNA binding specificity, gene regulation, binding motifs, high-throughput sequencing, position weight matrix, human genome, regulatory variants, protein-DNA interactions

Cite Scienmag News
APA MLA Chicago

Juliet Wilcox. (September 5, 2026). GHT-SELEX reveals strong sequence specificity in human transcription factors. Scienmag. https://scienmag.com/ght-selex-reveals-strong-sequence-specificity-in-human-transcription-factors/

Copy citation Download RIS

Tags: advances in DNA-binding assaysadvances in gene regulation researchcomplex DNA-binding behaviorDNA motif recognitiongene regulatory network analysisgene regulatory network mappinggenomic DNA-protein interactionsGHT-SELEX techniqueGHT-SELEX technologyhuman gene regulationhuman gene regulation mechanismshuman transcription factor diversityimpact on gene regulation researchmolecular biology of gene activationmolecular biology of gene expressiontranscription factor binding site complexitytranscription factor binding site mappingtranscription factor DNA-binding specificitytranscription factor intrinsic sequence preferencestranscription factor intrinsic specificitytranscription factor sequence preferencestranscription factor-DNA interaction analysis

Read Entire Article

         

        

Start the new Vibrations with a Medbed Franchise today!  

Protect your whole family with Quantum Orgo-Life® devices

  Advertising by Adpathway