This dataset contains the augmented ncRNA–ncRNA training data used in the GRIP-Transformer experiments. It includes 75,380 source–target RNA sequence pairs obtained through four-fold data augmentation of 18,845 non-augmented training interactions. Each record represents an interacting RNA pair used to train the encoder–decoder Transformer, where RNA1 is provided as the source sequence and RNA2 as the target sequence. The corresponding non-augmented training set is available in the GRIP-Transformer GitHub repository. The augmented file is deposited on Zenodo because its size prevents direct inclusion in the GitHub repository. The dataset was derived from RNA-KG data and subsequently prepared and augmented for the GRIP-Transformer experiments.
GRIP-Transformer ncRNA–ncRNA augmented training dataset / M. D'Ovidio, M.N.. - (2026). [10.5281/zenodo.22300231]
GRIP-Transformer ncRNA–ncRNA augmented training dataset
M. NicoliniSecondo
;E. CasiraghiPenultimo
;G. ValentiniUltimo
2026
Abstract
This dataset contains the augmented ncRNA–ncRNA training data used in the GRIP-Transformer experiments. It includes 75,380 source–target RNA sequence pairs obtained through four-fold data augmentation of 18,845 non-augmented training interactions. Each record represents an interacting RNA pair used to train the encoder–decoder Transformer, where RNA1 is provided as the source sequence and RNA2 as the target sequence. The corresponding non-augmented training set is available in the GRIP-Transformer GitHub repository. The augmented file is deposited on Zenodo because its size prevents direct inclusion in the GitHub repository. The dataset was derived from RNA-KG data and subsequently prepared and augmented for the GRIP-Transformer experiments.Pubblicazioni consigliate
I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.




