AIRBabel immunoglobulin and T-cell receptor allele sequences · leaders · RSSs

← back to search

RHESUS-VEQGXZP5P

AIRBabel's deterministic identifier, minted from the sequence (SPECIES-CODE+hash) i: a stable sequence-derived identifier, not an official allele name. This sequence's names from each source are listed under Allele names.

SpeciesMacaca mulatta LocusIGH TypeV FunctionalTrue Length290 nt
SpeciesMacaca mulatta (NCBITAXON:9544)
Locus / typeIGH / V
Functional iTrue
Length290 nt
Protein UID iRHESUS-VP-PCWLKC (shared by alleles with the same V-REGION amino-acid sequence)

Allele names i

Every name below denotes this exact sequence (RHESUS-VEQGXZP5P). None is canonical; they are labels, and the UID is the key.

Multiple allele designations. This one sequence is recorded under 2 different allele designations: IGHV3-12*01, LJI.Rh_IGHV3.8. These are identical normalized nucleotide sequences with different attributed names.

NameSchemeSource(s)
IGHV3-BTKS*03 IgLabel MUSA
IGHV3-OLUF*07 IgLabel OGRDB
IGHV3-12*01 IMGT IMGT
LJI.Rh_IGHV3.8 RhGLDB+ RhGLDB+

Other names & identifiers

Alternative and former names, e.g. IMGT clone names, plus accession identifiers.

NameKindSource(s)
BK063715.1 NCBI accession ncbi_nucleotide

External records i

NCBI nucleotide: link GenBank: GENBANK:BK063715 GenBank: GENBANK:BK063715.1 IMGT: IGHV3-12*01 OGRDB: IGHV3-OLUF*07

Literature references 6 total · 3 verified · 1 sequence-in-paper · 2 name-mentioned* i

PubMed publications linked to this allele. A sequence-in-paper match indicates that this allele's exact sequence was detected in the publication text or supplementary material; it is sequence-level evidence. Name-mentioned entries were auto-collected by name match and may refer to a different same-named allele; treat them as leads, not ground truth.

PMIDTitleEvidenceSource
31080066 verified MUSA
34875068 verified ncbi_nucleotide
35335026 verified ncbi_nucleotide
32866207 Mapping the immunogenic landscape of near-native HIV-1 envelope trimers in non-human primates. sequence in paper literature_supp
35419009 Addressing IGHV Gene Structural Diversity Enhances Immunoglobulin Repertoire Analysis: Lessons From Rhesus Macaque. name mentioned* literature_supp
38961347 Systematic characterization of immunoglobulin loci and deep sequencing of the expressed repertoire in the Atlantic cod (Gadus morhua). name mentioned* literature_supp

* matched by allele name only. The publication mentions this name, but AIRBabel cannot determine from the name alone whether it refers to this allele; the same name can denote a different allele in another species. This is useful for literature review but is not verified evidence.

Contributed by IMGT, MUSA, OGRDB, RhGLDB+, ncbi_nucleotide. See Sources & citations for each source's version, citation and licence.

RSS i

Recombination signal sequences flanking this allele. Click one to see every coding allele linked to that exact RSS variant.

PartSequenceSourceRSS UID
v_rsCACGGTGAGGGGAGGTCAGTGTGAGCCGACACAAACCTC MUSA RHESUS-RJNQKE
v_rsCACGGTGAGGGGAGGTCAGTGTGAGCCGACACAAACC IMGT RHESUS-RVBIXS

Leader

Signal-peptide coding sequence (spliced exon 1 + exon 2). The genomic intron between the two exons is deferred.

Leader (exon 1 + exon 2)SplitSourceLeader UID
ATGGAGTTGGGGCTGAGCTGGGTTTTCCTTGTTGCTATTTTAGAAGGTGTCCAGTGT 46 + 11 nt IMGT RHESUS-L2VK2X
ATGGAGTTGGGGCTGAGCTGGGTTTTCCTTGTTGCTATTTTAGAAGGTGTGTCCAGTGT 48 + 11 nt MUSA RHESUS-LCCXMJ

Find similar alleles

Ranks coding alleles by normalized edit-distance identity. Also available at /api/sequences/RHESUS-VEQGXZP5P/similar?min_identity=0.95&kind=nt.

Coding sequence (nt)

290 nt
GAGGTGCAGCTGGTGGAGTCTGGGGGAGGCTTGGTACAGCCTGGCCGGTCCCTGAGACCCTCCTGTGCAGCCTCTGGATTCACTTTCAGTAGCTATGGCATGCACTGGGTCCGCCAGGCTCCGGAAGAGGGGCTGGTGTGGGTTTCATACATTGGTAGTAGTACCATGTACTACGCAGACTCCGTGAAGGGCCGATTCACCATCTCCAGAGACAATGCCAAGAACTCGCTGTATCTGCAAATGAACAGCCTGAGAGCCGAGGACACGGCTGTGTATTACTGTGTGAGAGA

Amino-acid sequence (V-REGION)

96 aa
EVQLVESGGGLVQPGRSLRPSCAASGFTFSSYGMHWVRQAPEEGLVWVSYIGSSTMYYADSVKGRFTISRDNAKNSLYLQMNSLRAEDTAVYYCVR

Published by the source, and consistent with the nucleotide sequence above. i