• Home
  • Search
  • Predicting transcriptional activation domain function using Graph Neural Networks.
  • Open Access IconOpen Access
  • Cite Icon1
  • https://doi.org/10.1101/2024.05.08.593266Copy DOI Icon

Predicting transcriptional activation domain function using Graph Neural Networks.

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Analysis of factors that lead to the functionality of transcriptional activation domains remains a crucial and yet challenging task owing to the significant diversity in their sequences and their intrinsically disordered nature. Almost all existing methods that have aimed to predict activation domains have involved traditional machine learning approaches, such as logistic regression, that are unable to capture complex patterns in data or plain convolutional neural networks and have been limited in exploration of structural features. However, there is a tremendous potential in the inspection of the structural properties of activation domains, and an opportunity to investigate complex relationships between features of residues in the sequence. To address these, we have utilized the power of graph neural networks which can represent structural data in the form of nodes and edges, allowing nodes to exchange information among themselves. We have experimented with two kinds of graph formulations, one involving residues as nodes and the other assigning atoms to be the nodes. A logistic regression model was also developed to analyze feature importance. For all the models, several feature combinations were experimented with. The residue-level GNN model with amino acid type, residue position, acidic/basic/aromatic property and secondary structure feature combination gave the best performing model with accuracy, F1 score and AUROC of 97.9%, 71% and 97.1% respectively which outperformed other existing methods in the literature when applied on the dataset we used. Among the other structure-based features that were analyzed, the amphipathic property of helices also proved to be an important feature for classification. Logistic regression results showed that the most dominant feature that makes a sequence functional is the frequency of different types of amino acids in the sequence. Our results consistent have shown that functional sequences have more acidic and aromatic residues whereas basic residues are seen more in non-functional sequences.

Similar Papers
  • Discussion
  • Citations4

Monitoring Creutzfeldt-Jakob disease.

  • Aug 01, 1992
  • BMJ (Clinical research ed.)
  • D Lyons +1
  • Research Article
  • Citations232

Phosphorylation of Nrf2 in the transcription activation domain by casein kinase 2 (CK2) is critical for the nuclear translocation and transcription activation function of Nrf2 in IMR‐32 neuroblastoma cells

  • Feb 01, 2008
  • Journal of Biochemical and Molecular Toxicology
  • Patrick L Apopa +2
  • Research Article
  • Citations17

SUMO-Forest: A Cascade Forest based method for the prediction of SUMOylation sites on imbalanced data.

  • Mar 08, 2020
  • Gene
  • Ying Qian +3
  • Research Article
  • Citations231

Nature of Driving Force for Protein Folding: A Result From Analyzing the Statistical Potential

  • Jul 28, 1997
  • Physical Review Letters
  • Hao Li +2
  • Research Article

Amino‐Acid‐Encoded Enantioselective Photocatalysis in Self‐Assembled Porphyrin–Amino Acid Derivatives

  • Jun 02, 2025
  • European Journal of Organic Chemistry
  • Bowen Li +7
  • Research Article
  • Citations13

Amino Acid-Dependent Material Properties of Tetrapeptide Condensates

  • May 15, 2024
  • bioRxiv
  • Yi Zhang +4
  • Supplementary Content

Data mining for important amino acid residues in multiple sequence alignments and protein structures

  • Jan 01, 2014
  • University of Regensburg Publication Server (University of Regensburg)
  • Jan-Oliver Janda
  • Research Article
  • Citations2

Interpretable biophysical neural networks of transcriptional activation domains separate roles of protein abundance and coactivator binding

  • Sep 21, 2025
  • bioRxiv
  • Claire Leblanc +7
  • Research Article
  • Citations29

A Potential New Therapeutic Approach for Friedreich Ataxia: Induction of Frataxin Expression With TALE Proteins

  • Jan 01, 2013
  • Molecular Therapy - Nucleic Acids
  • Pierre Chapdelaine +4
  • Research Article
  • Citations26

Unleashing the power of SDN and GNN for network anomaly detection: State‐of‐the‐art, challenges, and future directions

  • Jul 30, 2023
  • SECURITY AND PRIVACY
  • Archan Dhadhania +5
  • Research Article
  • Citations196

A High-Throughput Mutational Scan of an Intrinsically Disordered Acidic Transcriptional Activation Domain.

  • Mar 07, 2018
  • Cell systems
  • Max V Staller +5
  • PDF
  • Research Article
  • Citations60

Nucleosome Transactions on the Promoters of the YeastGAL and PHO Genes

  • Oct 01, 1997
  • Journal of Biological Chemistry
  • D Lohr
  • PDF
  • Research Article
  • Citations43

Critical Amino Acids in the Transcriptional Activation Domain of the Herpesvirus Protein VP16 Are Solvent-exposed in Highly Mobile Protein Segments: AN INTRINSIC FLUORESCENCE STUDY

  • Mar 01, 1996
  • Journal of Biological Chemistry
  • Fan Shen +4
  • Research Article

Federated Learning-Based Distributed Autoencoder for Industrial Big Data Anomaly Detection: Integrating LSTM, GRU, and CNN Models

  • Jul 21, 2025
  • Informatica
  • Xiaoli Li +1
  • Research Article

Imprint of the Forgotten: Stealthy Membership Inference in Unlearned Graph Neural Networks

  • Mar 14, 2026
  • He Zhang +4
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.