Data, Results, and Scripts for "Alignment of RNA Secondary Structures with Arbitrary Pseudoknots using Structural Sequences"

Fuente: Zenodo
Saved in:
Bibliographic Details
Main Authors: Tesei, Luca, LEVI, FRANCESCA, QUADRINI, Michela, Merelli, Emanuela
Format: Recurso digital
Published: Zenodo 2026
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866902337467973632
author Tesei, Luca
LEVI, FRANCESCA
QUADRINI, Michela
Merelli, Emanuela
author_facet Tesei, Luca
LEVI, FRANCESCA
QUADRINI, Michela
Merelli, Emanuela
contents <p><span lang="EN-US">This repository contains datasets, results, and scripts associated with the following manuscript:</span></p> <p><span lang="EN-GB">Tesei, L., Levi, F., Quadrini, M., Merelli, E. (2026) </span><em><span lang="EN-US">“Alignment of RNA Secondary Structures with Arbitrary Pseudoknots using Structural Sequences”</span></em></p> <p><span lang="EN-US">The software tool SERNAlign, used to process the RNA secondary structures and generate </span><span lang="EN-US">SERNA</span><span lang="EN-US"> and </span><span lang="EN-US">SERNA-NC</span><span lang="EN-US"> distances is available:</span></p> <p><a href="https://github.com/bdslab/sernalign"><span lang="EN-US">https://github.com/bdslab/sernalign</span></a><span> </span></p> <p><span lang="EN-US">Please refer to the documentation of the SERNALign tool at:</span></p> <p><a href="https://github.com/bdslab/sernalign/blob/master/README.md"><span lang="EN-US">https://github.com/bdslab/sernalign/blob/master/README.md</span></a><span> </span></p> <p><span lang="EN-US">This repository includes curated collections of RNA structures (with and without pseudoknots), comparative analyses performed with a wide set of structural alignment tools, and scripts for clustering analysis. The following distances were used to generate the comparative results included in this repository:</span><span lang="EN-US"> SERNA</span><span lang="EN-US">, </span><span lang="EN-US">SERNA-NC</span><span lang="EN-US">, </span><span lang="EN-US">ASPRA</span><span lang="EN-US">, </span><span lang="EN-US">PSMAlign</span><span lang="EN-US">, </span><span lang="EN-US">RAG-2D</span><span lang="EN-US">, </span><span lang="EN-US">Genus</span><span lang="EN-US">, and</span><span lang="EN-US"> PskOrder</span><span lang="EN-US">. </span><span lang="EN-US">For completeness, the tables also include metrics computed using simple global descriptors, namely the number of base pairs (</span><span lang="EN-US">BP</span><span lang="EN-US">), number of GC nucleotides in the sequence (</span><span lang="EN-US">GC-Seq</span><span lang="EN-US">), number of G–C base pairs (</span><span lang="EN-US">GC-Pairs</span><span lang="EN-US">), and sequence length (</span><span lang="EN-US">Length</span><span lang="EN-US">).</span></p> <p><span lang="EN-US">ASPRA</span><span lang="EN-US"> is the distance obtained by aligning Structural RNA Trees of RNA secondary structures with arbitrary pseudoknots [1]. </span><span lang="EN-US">BP</span><span lang="EN-US"> and </span><span lang="EN-US">Length</span><span lang="EN-US"> represent the distances computed from the number of base pairs and from the sequence length of each molecule, respectively. </span><span lang="EN-US">GC-Seq</span><span lang="EN-US"> and </span><span lang="EN-US">GC-Pairs</span><span lang="EN-US"> are distances based on the number of occurrences of nucleotides G and C, and on the number of G–C and C–G base pairs, respectively. </span><span lang="EN-US">SERNA</span><span lang="EN-US"> and </span><span lang="EN-US">SERNA-NC</span><span lang="EN-US"> represent the distances computed by aligning Structural Sequences, with and without structural constraints. </span><span lang="EN-US">Genus</span><span lang="EN-US"> and </span><span lang="EN-US">PskOrder</span><span lang="EN-US"> correspond to distances computed from the genus of each RNA molecule [2] and from the pseudoknot order [3]. </span><span lang="EN-US">PSMAlign</span><span lang="EN-US"> and </span><span lang="EN-US">RAG-2D</span><span lang="EN-US"> are the distances produced by the homonymous algorithms [4] and [5].</span></p> <p><strong><span lang="EN-US">Pseudoknots_Dataset</span></strong><strong><span lang="EN-US"> Folder</span></strong></p> <p><span lang="EN-US">This folder contains RNA molecules featuring pseudoknotted secondary structures, extracted from the <strong>Pseudobase++</strong> database. The dataset is organized into two main subfolders: <em>pseudoknots_3class</em> and <em>pseudoknots_8class</em>. The first dataset groups molecules into three pseudoknot classes (H, HLout, and LL_HHH), while the second organizes them into eight classes (H, HHH, HLin, HLout, HLout_HHH, LL, LL_HHH, and HLout_HLin), providing a finer level of structural categorization. Within each dataset, the molecules are organized into subdirectories corresponding to different structural formats (BPSEQ, CT, and DB), each representing an alternative encoding of the RNA secondary structure. These formats allow the pseudoknotted structures to be processed by different comparison and analysis tools used in this study.</span></p> <p><strong><span lang="EN-US">Phylogenetic_Dataset</span></strong><strong><span lang="EN-US"> Folder</span></strong></p> <p><span lang="EN-US">This folder contains ribosomal RNA molecules extracted from the CRW2 (Comparative RNA Web-2) database [6] and organized into three top-level subfolders corresponding to the three domains of life: Archaea, Bacteria, and Eukaryota. Each of these domain-specific folders includes three additional subdirectories, one for each ribosomal RNA type present in the dataset, namely 5S, 16S, and 23S. Within each RNA-type folder, the molecules are provided in several structural formats (Bpseq, CT, and DB), which represent different encodings of RNA secondary structures, with or without non-canonical interactions. This organization ensures compatibility with the various alignment and comparison tools employed in our analyses.</span></p> <p><strong><span lang="EN-US">Pseudoknots_Dataset_Distances_And_Labels</span></strong><strong><span lang="EN-US"> Folder</span></strong></p> <p><span lang="EN-US">This folder contains the pseudoknot datasets and is organized into three subfolders: <em>pseudoknots_3class</em>, <em>pseudoknots_8class</em>, and <em>Labels</em>. The first two subfolders collect the outputs produced by the different alignment and comparison tools applied to the molecules belonging to each pseudoknot dataset. All result files follow a uniform naming convention of the form </span><code><span lang="EN-US">ToolName_DatasetName.csv</span></code><span lang="EN-US">, ensuring that each result can be directly associated with both the tool used and the dataset to which it refers.</span></p> <p><span lang="EN-US">The Labels folder contains the annotation files for both pseudoknot datasets. Each file follows the naming convention </span><code><span lang="EN-US">Labels_DatasetName_pseudoknotType.csv</span></code><span lang="EN-US">, where the dataset name identifies whether the file refers to the three-class or eight-class dataset, and <em>pseudoknotType</em> denotes the structural pseudoknot class assigned to each molecule. In the three-class dataset, pseudoknots are classified into H, HLout, and LL_HHH. In the eight-class dataset, the classification is refined into H, HHH, HLin, HLout, HLout_HHH, LL, LL_HHH, and HLout_HLin. These label files provide the necessary information to associate every molecule with its corresponding pseudoknot class and support downstream analyses based on structural categories.</span></p> <p><strong><span lang="EN-GB">Phylogenetic_Dataset_Distances_And_Labels</span></strong><strong><span lang="EN-GB"> Folder</span></strong></p> <p><span lang="EN-US">This folder contains the phylogenetic datasets and is structured into four subfolders: Archaea, Bacteria, Eukaryota, and Labels. The first three folders collect the outputs produced by the different alignment and comparison tools applied to the molecules extracted from each domain of life. For each tool and domain, the results are distinguished by RNA type, and all files follow a consistent naming convention of the form </span><code><span lang="EN-US">ToolName_Domain_RNAType.csv</span></code><span lang="EN-US">, where the domain is one among Archaea, Bacteria, or Eukaryota, and the RNA type is one among 5S, 16S, or 23S. This convention guarantees that each file can be unambiguously associated with the corresponding tool, domain, and ribosomal RNA type.</span></p> <p><span lang="EN-US">The Labels folder contains all annotation files for the phylogenetic datasets. Each file follows the naming convention </span><code><span lang="EN-US">Labels_Domain_Type_taxonomicClassification.csv</span></code><span lang="EN-US">, where the domain indicates Archaea, Bacteria, or Eukaryota, the type specifies whether the file refers to 5S, 16S, or 23S molecules, and <em>taxonomicClassification</em> denotes that the file includes taxonomic metadata associated with each molecule, namely its phylum, order, and class.</span></p> <p><strong><span lang="EN-US">Scripts</span></strong><strong><span lang="EN-US"> Folder</span></strong></p> <p><span lang="EN-US">This folder contains the Python scripts to perform the hierarchical clustering of all datasets using the different linkages and evaluating the obtained clusters with the metrics. The script </span><span lang="EN-US">ClusterMatrix.py</span><span lang="EN-US"> can be used for csv with distances while the script </span><span lang="EN-US">ClusterFeatures.py</span><span lang="EN-US"> can be used for csv with features (RAG-2D).</span></p> <p><strong><span lang="EN-US">Numerical_Results</span></strong><strong><span lang="EN-US"> Folder</span></strong></p> <p><span lang="EN-US">This folder contains all the tables with numerical values of the metrics for each dataset and each distance, computed by the provided scripts. </span></p> <p><strong><span lang="EN-GB">References</span></strong></p> <p><span>[1] Quadrini, M., Tesei, L., and Merelli, E. (2020). </span><span lang="EN-US">ASPRAlign: a tool for the alignment of RNA secondary structures with arbitrary pseudoknots. Bioinformatics, 36(11), 3578-3579. </span></p> <p><span lang="EN-US">[2] Andersen, J. E., Penner, R. C., Reidys, C. M., and Waterman, M. S. (2013). Topological classification and enumeration of RNA structures by genus. Journal of mathematical biology, 67(5), 1261-1278.</span></p> <p><span lang="EN-US">[3] Zok, T., Badura, J., Swat, S., Figurski, K., Popenda, M., and Antczak, M. (2020). New models and algorithms for RNA pseudoknot order assignment. International Journal of Applied Mathematics and Computer Science, 30(2), 315-324.</span></p> <p><span lang="EN-US">[4] Chiu, J. K. H., and Chen, Y. P. P. (2015). Pairwise RNA secondary structure alignment with conserved stem pattern. Bioinformatics, 31(24), 3914-3921.</span></p> <p><span lang="EN-US">[5] Gan, H. H., Fera, D., Zorn, J., Shiffeldrim, N., Tang, M., Laserson, U., Kim N. and Schlick, T. (1987). RAG: RNA-As-Graphs database—concepts, analysis, and features. Nutrition and Health, 5(1-2), 1285-1291.</span></p> <p><span lang="EN-US">[6] Cannone, J. J., Subramanian, S., Schnare, M. N., Collett, J. R., D'Souza, L. M., Du, Y., Feng B., Lin N., </span><a href="https://link.springer.com/article/10.1186/1471-2105-3-2#auth-Lakshmi_V-Madabusi-Aff1-Aff4"><span lang="EN-US">Lakshmi V., Madabusi</span></a><span lang="EN-US">, </span><a href="https://link.springer.com/article/10.1186/1471-2105-3-2#auth-Kirsten_M-M_ller-Aff1-Aff5"><span lang="EN-US">Müller</span></a><span lang="EN-US"> K. M.,  </span><a href="https://link.springer.com/article/10.1186/1471-2105-3-2#auth-Nupur-Pande-Aff1"><span lang="EN-US">Pande</span></a><span lang="EN-US"> N., <a href="https://link.springer.com/article/10.1186/1471-2105-3-2#auth-Zhidi-Shang-Aff1"><span>Zhidi Shang</span></a> Z., </span><a href="https://link.springer.com/article/10.1186/1471-2105-3-2#auth-Nan-Yu-Aff1"><span lang="EN-US">Yu</span></a><span lang="EN-US"> N. and Gutell, R. R. (2002). The comparative RNA web (CRW) site: an online database of comparative sequence and structure information for ribosomal, intron, and other RNAs. BMC Bioinformatics, 3(1), 2.</span></p> <p><span lang="EN-US">[7] Taufer, M., Licon, A., Araiza, R., Mireles, D., Van Batenburg, F. H. D., Gultyaev, A. P., & Leung, M. Y. (2009). PseudoBase++: an extension of PseudoBase for easy searching, formatting and visualization of pseudoknots. Nucleic acids research, 37(suppl_1), D127-D135.</span></p> <p><span lang="EN-US"><strong><span lang="EN-GB">Funding</span></strong><br>This work was supported by the European Union - NextGenerationEU - National Recovery and Resilience Plan (NRRP), Mission 4, Component 2, Investment 1.1, under the PRIN 2022 PNRR call (Min. Decree No. 1409, dated September 14, 2022), project: P2022FFEWN RNA secondary structures and their relationship with function: application to non-coding RNAs (RNA2Fun), CUP: J53D23014960001.</span></p>
format Recurso digital
id zenodo_https___doi_org_10_5281_zenodo_18292182
institution Zenodo
language
publishDate 2026
publisher Zenodo
record_format zenodo
spellingShingle Data, Results, and Scripts for "Alignment of RNA Secondary Structures with Arbitrary Pseudoknots using Structural Sequences"
Tesei, Luca
LEVI, FRANCESCA
QUADRINI, Michela
Merelli, Emanuela
<p><span lang="EN-US">This repository contains datasets, results, and scripts associated with the following manuscript:</span></p> <p><span lang="EN-GB">Tesei, L., Levi, F., Quadrini, M., Merelli, E. (2026) </span><em><span lang="EN-US">“Alignment of RNA Secondary Structures with Arbitrary Pseudoknots using Structural Sequences”</span></em></p> <p><span lang="EN-US">The software tool SERNAlign, used to process the RNA secondary structures and generate </span><span lang="EN-US">SERNA</span><span lang="EN-US"> and </span><span lang="EN-US">SERNA-NC</span><span lang="EN-US"> distances is available:</span></p> <p><a href="https://github.com/bdslab/sernalign"><span lang="EN-US">https://github.com/bdslab/sernalign</span></a><span> </span></p> <p><span lang="EN-US">Please refer to the documentation of the SERNALign tool at:</span></p> <p><a href="https://github.com/bdslab/sernalign/blob/master/README.md"><span lang="EN-US">https://github.com/bdslab/sernalign/blob/master/README.md</span></a><span> </span></p> <p><span lang="EN-US">This repository includes curated collections of RNA structures (with and without pseudoknots), comparative analyses performed with a wide set of structural alignment tools, and scripts for clustering analysis. The following distances were used to generate the comparative results included in this repository:</span><span lang="EN-US"> SERNA</span><span lang="EN-US">, </span><span lang="EN-US">SERNA-NC</span><span lang="EN-US">, </span><span lang="EN-US">ASPRA</span><span lang="EN-US">, </span><span lang="EN-US">PSMAlign</span><span lang="EN-US">, </span><span lang="EN-US">RAG-2D</span><span lang="EN-US">, </span><span lang="EN-US">Genus</span><span lang="EN-US">, and</span><span lang="EN-US"> PskOrder</span><span lang="EN-US">. </span><span lang="EN-US">For completeness, the tables also include metrics computed using simple global descriptors, namely the number of base pairs (</span><span lang="EN-US">BP</span><span lang="EN-US">), number of GC nucleotides in the sequence (</span><span lang="EN-US">GC-Seq</span><span lang="EN-US">), number of G–C base pairs (</span><span lang="EN-US">GC-Pairs</span><span lang="EN-US">), and sequence length (</span><span lang="EN-US">Length</span><span lang="EN-US">).</span></p> <p><span lang="EN-US">ASPRA</span><span lang="EN-US"> is the distance obtained by aligning Structural RNA Trees of RNA secondary structures with arbitrary pseudoknots [1]. </span><span lang="EN-US">BP</span><span lang="EN-US"> and </span><span lang="EN-US">Length</span><span lang="EN-US"> represent the distances computed from the number of base pairs and from the sequence length of each molecule, respectively. </span><span lang="EN-US">GC-Seq</span><span lang="EN-US"> and </span><span lang="EN-US">GC-Pairs</span><span lang="EN-US"> are distances based on the number of occurrences of nucleotides G and C, and on the number of G–C and C–G base pairs, respectively. </span><span lang="EN-US">SERNA</span><span lang="EN-US"> and </span><span lang="EN-US">SERNA-NC</span><span lang="EN-US"> represent the distances computed by aligning Structural Sequences, with and without structural constraints. </span><span lang="EN-US">Genus</span><span lang="EN-US"> and </span><span lang="EN-US">PskOrder</span><span lang="EN-US"> correspond to distances computed from the genus of each RNA molecule [2] and from the pseudoknot order [3]. </span><span lang="EN-US">PSMAlign</span><span lang="EN-US"> and </span><span lang="EN-US">RAG-2D</span><span lang="EN-US"> are the distances produced by the homonymous algorithms [4] and [5].</span></p> <p><strong><span lang="EN-US">Pseudoknots_Dataset</span></strong><strong><span lang="EN-US"> Folder</span></strong></p> <p><span lang="EN-US">This folder contains RNA molecules featuring pseudoknotted secondary structures, extracted from the <strong>Pseudobase++</strong> database. The dataset is organized into two main subfolders: <em>pseudoknots_3class</em> and <em>pseudoknots_8class</em>. The first dataset groups molecules into three pseudoknot classes (H, HLout, and LL_HHH), while the second organizes them into eight classes (H, HHH, HLin, HLout, HLout_HHH, LL, LL_HHH, and HLout_HLin), providing a finer level of structural categorization. Within each dataset, the molecules are organized into subdirectories corresponding to different structural formats (BPSEQ, CT, and DB), each representing an alternative encoding of the RNA secondary structure. These formats allow the pseudoknotted structures to be processed by different comparison and analysis tools used in this study.</span></p> <p><strong><span lang="EN-US">Phylogenetic_Dataset</span></strong><strong><span lang="EN-US"> Folder</span></strong></p> <p><span lang="EN-US">This folder contains ribosomal RNA molecules extracted from the CRW2 (Comparative RNA Web-2) database [6] and organized into three top-level subfolders corresponding to the three domains of life: Archaea, Bacteria, and Eukaryota. Each of these domain-specific folders includes three additional subdirectories, one for each ribosomal RNA type present in the dataset, namely 5S, 16S, and 23S. Within each RNA-type folder, the molecules are provided in several structural formats (Bpseq, CT, and DB), which represent different encodings of RNA secondary structures, with or without non-canonical interactions. This organization ensures compatibility with the various alignment and comparison tools employed in our analyses.</span></p> <p><strong><span lang="EN-US">Pseudoknots_Dataset_Distances_And_Labels</span></strong><strong><span lang="EN-US"> Folder</span></strong></p> <p><span lang="EN-US">This folder contains the pseudoknot datasets and is organized into three subfolders: <em>pseudoknots_3class</em>, <em>pseudoknots_8class</em>, and <em>Labels</em>. The first two subfolders collect the outputs produced by the different alignment and comparison tools applied to the molecules belonging to each pseudoknot dataset. All result files follow a uniform naming convention of the form </span><code><span lang="EN-US">ToolName_DatasetName.csv</span></code><span lang="EN-US">, ensuring that each result can be directly associated with both the tool used and the dataset to which it refers.</span></p> <p><span lang="EN-US">The Labels folder contains the annotation files for both pseudoknot datasets. Each file follows the naming convention </span><code><span lang="EN-US">Labels_DatasetName_pseudoknotType.csv</span></code><span lang="EN-US">, where the dataset name identifies whether the file refers to the three-class or eight-class dataset, and <em>pseudoknotType</em> denotes the structural pseudoknot class assigned to each molecule. In the three-class dataset, pseudoknots are classified into H, HLout, and LL_HHH. In the eight-class dataset, the classification is refined into H, HHH, HLin, HLout, HLout_HHH, LL, LL_HHH, and HLout_HLin. These label files provide the necessary information to associate every molecule with its corresponding pseudoknot class and support downstream analyses based on structural categories.</span></p> <p><strong><span lang="EN-GB">Phylogenetic_Dataset_Distances_And_Labels</span></strong><strong><span lang="EN-GB"> Folder</span></strong></p> <p><span lang="EN-US">This folder contains the phylogenetic datasets and is structured into four subfolders: Archaea, Bacteria, Eukaryota, and Labels. The first three folders collect the outputs produced by the different alignment and comparison tools applied to the molecules extracted from each domain of life. For each tool and domain, the results are distinguished by RNA type, and all files follow a consistent naming convention of the form </span><code><span lang="EN-US">ToolName_Domain_RNAType.csv</span></code><span lang="EN-US">, where the domain is one among Archaea, Bacteria, or Eukaryota, and the RNA type is one among 5S, 16S, or 23S. This convention guarantees that each file can be unambiguously associated with the corresponding tool, domain, and ribosomal RNA type.</span></p> <p><span lang="EN-US">The Labels folder contains all annotation files for the phylogenetic datasets. Each file follows the naming convention </span><code><span lang="EN-US">Labels_Domain_Type_taxonomicClassification.csv</span></code><span lang="EN-US">, where the domain indicates Archaea, Bacteria, or Eukaryota, the type specifies whether the file refers to 5S, 16S, or 23S molecules, and <em>taxonomicClassification</em> denotes that the file includes taxonomic metadata associated with each molecule, namely its phylum, order, and class.</span></p> <p><strong><span lang="EN-US">Scripts</span></strong><strong><span lang="EN-US"> Folder</span></strong></p> <p><span lang="EN-US">This folder contains the Python scripts to perform the hierarchical clustering of all datasets using the different linkages and evaluating the obtained clusters with the metrics. The script </span><span lang="EN-US">ClusterMatrix.py</span><span lang="EN-US"> can be used for csv with distances while the script </span><span lang="EN-US">ClusterFeatures.py</span><span lang="EN-US"> can be used for csv with features (RAG-2D).</span></p> <p><strong><span lang="EN-US">Numerical_Results</span></strong><strong><span lang="EN-US"> Folder</span></strong></p> <p><span lang="EN-US">This folder contains all the tables with numerical values of the metrics for each dataset and each distance, computed by the provided scripts. </span></p> <p><strong><span lang="EN-GB">References</span></strong></p> <p><span>[1] Quadrini, M., Tesei, L., and Merelli, E. (2020). </span><span lang="EN-US">ASPRAlign: a tool for the alignment of RNA secondary structures with arbitrary pseudoknots. Bioinformatics, 36(11), 3578-3579. </span></p> <p><span lang="EN-US">[2] Andersen, J. E., Penner, R. C., Reidys, C. M., and Waterman, M. S. (2013). Topological classification and enumeration of RNA structures by genus. Journal of mathematical biology, 67(5), 1261-1278.</span></p> <p><span lang="EN-US">[3] Zok, T., Badura, J., Swat, S., Figurski, K., Popenda, M., and Antczak, M. (2020). New models and algorithms for RNA pseudoknot order assignment. International Journal of Applied Mathematics and Computer Science, 30(2), 315-324.</span></p> <p><span lang="EN-US">[4] Chiu, J. K. H., and Chen, Y. P. P. (2015). Pairwise RNA secondary structure alignment with conserved stem pattern. Bioinformatics, 31(24), 3914-3921.</span></p> <p><span lang="EN-US">[5] Gan, H. H., Fera, D., Zorn, J., Shiffeldrim, N., Tang, M., Laserson, U., Kim N. and Schlick, T. (1987). RAG: RNA-As-Graphs database—concepts, analysis, and features. Nutrition and Health, 5(1-2), 1285-1291.</span></p> <p><span lang="EN-US">[6] Cannone, J. J., Subramanian, S., Schnare, M. N., Collett, J. R., D'Souza, L. M., Du, Y., Feng B., Lin N., </span><a href="https://link.springer.com/article/10.1186/1471-2105-3-2#auth-Lakshmi_V-Madabusi-Aff1-Aff4"><span lang="EN-US">Lakshmi V., Madabusi</span></a><span lang="EN-US">, </span><a href="https://link.springer.com/article/10.1186/1471-2105-3-2#auth-Kirsten_M-M_ller-Aff1-Aff5"><span lang="EN-US">Müller</span></a><span lang="EN-US"> K. M.,  </span><a href="https://link.springer.com/article/10.1186/1471-2105-3-2#auth-Nupur-Pande-Aff1"><span lang="EN-US">Pande</span></a><span lang="EN-US"> N., <a href="https://link.springer.com/article/10.1186/1471-2105-3-2#auth-Zhidi-Shang-Aff1"><span>Zhidi Shang</span></a> Z., </span><a href="https://link.springer.com/article/10.1186/1471-2105-3-2#auth-Nan-Yu-Aff1"><span lang="EN-US">Yu</span></a><span lang="EN-US"> N. and Gutell, R. R. (2002). The comparative RNA web (CRW) site: an online database of comparative sequence and structure information for ribosomal, intron, and other RNAs. BMC Bioinformatics, 3(1), 2.</span></p> <p><span lang="EN-US">[7] Taufer, M., Licon, A., Araiza, R., Mireles, D., Van Batenburg, F. H. D., Gultyaev, A. P., & Leung, M. Y. (2009). PseudoBase++: an extension of PseudoBase for easy searching, formatting and visualization of pseudoknots. Nucleic acids research, 37(suppl_1), D127-D135.</span></p> <p><span lang="EN-US"><strong><span lang="EN-GB">Funding</span></strong><br>This work was supported by the European Union - NextGenerationEU - National Recovery and Resilience Plan (NRRP), Mission 4, Component 2, Investment 1.1, under the PRIN 2022 PNRR call (Min. Decree No. 1409, dated September 14, 2022), project: P2022FFEWN RNA secondary structures and their relationship with function: application to non-coding RNAs (RNA2Fun), CUP: J53D23014960001.</span></p>
title Data, Results, and Scripts for "Alignment of RNA Secondary Structures with Arbitrary Pseudoknots using Structural Sequences"
url https://doi.org/10.5281/zenodo.18292182