Skip to content
Draft
Show file tree
Hide file tree
Changes from all commits
Commits
File filter

Filter by extension

Filter by extension

Conversations
Failed to load comments.
Loading
Jump to
Jump to file
Failed to load files.
Loading
Diff view
Diff view
3 changes: 2 additions & 1 deletion data/enpen/collection.json
Original file line number Diff line number Diff line change
Expand Up @@ -24,6 +24,7 @@
},
"dataset_order": [
"enpen/enterovirus/ev-d68",
"enpen/enterovirus/cva16"
"enpen/enterovirus/cva16",
"enpen/enterovirus/cva10"
]
}
3 changes: 3 additions & 0 deletions data/enpen/enterovirus/cva10/CHANGELOG.md
Original file line number Diff line number Diff line change
@@ -0,0 +1,3 @@
## Unreleased

Initial release
68 changes: 68 additions & 0 deletions data/enpen/enterovirus/cva10/README.md
Original file line number Diff line number Diff line change
@@ -0,0 +1,68 @@
# Coxsackievirus A10 dataset

| Key | Value |
|----------------------|-----------------------------------------------------------------------|
| authors | [Alejandra González-Sánchez](https://www.vallhebron.com/en/professionals/alejandra-gonzalez-sanchez), [Nadia Neuner-Jehle](https://eve-lab.org/people/nadia-neuner), [Emma B. Hodcroft](https://eve-lab.org/people/emma-hodcroft/), [ENPEN](https://escv.eu/european-non-polio-enterovirus-network-enpen/) |
| name | Coxsackievirus A10 |
| reference | [AY421767.1](https://www.ncbi.nlm.nih.gov/nuccore/AY421767.1) |
| workflow | <https://github.com/enterovirus-phylo/nextclade_a10> |
| path | `enpen/enterovirus/cva10` |
| clade definitions | A-H |

## Scope of this dataset

This dataset uses the [Static Inferred Ancestor](https://github.com/enterovirus-phylo/nextclade_a10/blob/master/resources/inferred-root.fasta) instead of the historical Kowalik prototype sequence ([AY421767.1](https://www.ncbi.nlm.nih.gov/nuccore/AY421767.1)). It is intended for broad subgenogroup classification, mutation quality control, and phylogenetic analysis of CVA10 diversity.

***Note:** Kowalik reference sequence is substantially diverged from currently circulating strains. This is common for many enterovirus datasets, in contrast to some other virus datasets (e.g., seasonal influenza) where the reference is updated more frequently to match recent sequences.*

To address this, the dataset is *rooted* on a Static Inferred Ancestor, a phylogenetically reconstructed ancestral sequence near the tree root. This provides a stable reference point that can be used as an alternative for mutation calling.

## Features

This dataset supports:

- Assignment of subgenotypes
- Phylogenetic placement
- Sequence quality control (QC)

## Subgenogroups of Coxsackievirus A10

Coxsackievirus A10 is divided into subgenogroups A, B, C, D, E (VP1 only), F, G and H (VP1 only).

***Note:** Genotypes E and H are based on VP1 sequences only.*

Overall, these designations are based on phylogenetic structure and characteristic mutations, and are widely used in molecular epidemiology, similar to subgenotype systems for other enteroviruses. Unlike influenza (H1N1, H3N2) or SARS-CoV-2, there is no universally standardized global lineage nomenclature for enteroviruses; naming instead follows conventions established in published studies and surveillance practices.

## Related Enteroviruses

CVA10 is closely related to other EV-A viruses, including CVA8, CVA16, and EV-A71. If you are not certain that your sequences contain only CVA10, we recommend using the "[Multiple Datasets](https://docs.nextstrain.org/projects/nextclade/en/stable/user/nextclade-web/getting-started.html#multi-dataset-mode)" tab instead of "Single Dataset".

This prevents Nextclade from forcing sequences to align to the CVA10 reference tree. For example, EV-A71 sequences may still align and receive a clade assignment (often near recombinant forms).

Please be cautious when working with short genes or fragments (e.g., 5'UTR sequences). These regions can be highly conserved across EV-A viruses, making genogroup and subgenogroup assignment prone to errors. In addition, such fragments may originate from recombinant genomes. Recombination is common in enteroviruses, and when analyzing only a fragment, this may go undetected.

If you are unsure how to proceed, please contact us. We are happy to assist.

## Reference types

This dataset includes several reference points used in analyses:

- *Static Inferred Ancestor:* Reconstructed ancestral sequence inferred with an outgroup, representing the likely founder of CVA10. Serves as a stable reference.

- *Parent:* The nearest ancestral node of a sample in the tree, used to infer branch-specific mutations.

- *Clade founder:* The inferred ancestral node defining a clade (e.g., A, B). Mutations "since clade founder" describe changes that define that clade.

- *Reference:* RefSeq or similarly established prototype sequence. Here Kowalik (AY421767.1).

- *Tree root:* Corresponds to the root of the tree, it may change in future updates as more data become available.

All references use the coordinate system of the Kowalik sequence.

## Issues & Contact

- For questions or suggestions, please [open an issue](https://github.com/enterovirus-phylo/nextclade_a10/issues) or email: eve-group[at]swisstph.ch

## What is a Nextclade dataset?

A Nextclade dataset includes the reference sequence, genome annotations, tree, clade definitions, and QC rules. Learn more in the [Nextclade documentation](https://docs.nextstrain.org/projects/nextclade/en/stable/user/datasets.html).
17 changes: 17 additions & 0 deletions data/enpen/enterovirus/cva10/genome_annotation.gff3
Original file line number Diff line number Diff line change
@@ -0,0 +1,17 @@
##gff-version 3
#!gff-spec-version 1.21
#!processor NCBI annotwriter
##sequence-region AY421767.1 1 7409
##species https://www.ncbi.nlm.nih.gov/Taxonomy/Browser/wwwtax.cgi?id=42769
AY421767.1 Genbank region 1 7409 . + . ID=AY421767.1:1..7409;Dbxref=taxon:42769;gb-acronym=CV-A10;gbkey=Src;isolate=CA10;mol_type=genomic RNA;strain=Kowalik
AY421767.1 Genbank CDS 745 951 . + . Name=VP4;product=VP4;gbkey=Prot;ID=id-AAR38847.1:1..69
AY421767.1 Genbank CDS 952 1716 . + . Name=VP2;product=VP2;gbkey=Prot;ID=id-AAR38847.1:70..324
AY421767.1 Genbank CDS 1717 2436 . + . Name=VP3;product=VP3;gbkey=Prot;ID=id-AAR38847.1:325..564
AY421767.1 Genbank CDS 2437 3330 . + . Name=VP1;product=VP1;gbkey=Prot;ID=id-AAR38847.1:565..862
AY421767.1 Genbank CDS 3331 3780 . + . Name=2A;product=2A;gbkey=Prot;ID=id-AAR38847.1:863..1012
AY421767.1 Genbank CDS 3781 4077 . + . Name=2B;product=2B;gbkey=Prot;ID=id-AAR38847.1:1013..1111
AY421767.1 Genbank CDS 4078 5064 . + . Name=2C;product=2C;gbkey=Prot;ID=id-AAR38847.1:1112..1440
AY421767.1 Genbank CDS 5065 5322 . + . Name=3A;product=3A;gbkey=Prot;ID=id-AAR38847.1:1441..1526
AY421767.1 Genbank CDS 5323 5388 . + . Name=3B;product=3B;gbkey=Prot;ID=id-AAR38847.1:1527..1548
AY421767.1 Genbank CDS 5389 5937 . + . Name=3C;product=3C;gbkey=Prot;ID=id-AAR38847.1:1549..1731
AY421767.1 Genbank CDS 5938 7323 . + . Name=3D;product=3D;gbkey=Prot;ID=id-AAR38847.1:1732..2193
Loading