Yum, tasty mutations...

MutationTaster - study a chromosomal position

MTQE documentation
NEVER press reload or F5 - unless you want to start from the very beginning.
input seems to be ok - now mapping the variant to the different transcripts...
found 1 transcript(s)...
Querying Taster for transcript #1: ENST00000225964
MT speed 0 s - this script 3.4 s

Results


genesymbolpredictionprobabilitymodelprediction
problem
splicingClinVaramino acid changesvariant typedbSNP IDprotein lengthfile
COL1A1disease_causing_automatic0.99999999584047simple_aaeaffected0G839Ssingle base exchangers72653131show file

Taster files

Yum, tasty mutations...

mutation t@sting

documentation

Prediction

disease causing

Model: simple_aae, prob: 0.99999999584047 (classification due to ClinVar, real probability is shown anyway)      (explain)
Summary
  • amino acid sequence changed
  • known disease mutation at this position (HGMD CM960323)
  • known disease mutation: rs17331 (pathogenic)
  • protein features (might be) affected
  • splice site changes
hyperlink
analysed issue analysis result
name of alteration no title
alteration (phys. location) chr17:48267406C>TN/A show variant in all transcripts   IGV
HGNC symbol COL1A1
Ensembl transcript ID ENST00000225964
Genbank transcript ID NM_000088
UniProt peptide P02452
alteration type single base exchange
alteration region CDS
DNA changes c.2515G>A
cDNA.2634G>A
g.11588G>A
AA changes G839S Score: 56 explain score(s)
position(s) of altered AA
if AA alteration in CDS
839
frameshift no
known variant Reference ID: rs72653131
Allele 'T' was neither found in ExAC nor 1000G.
known disease mutation: rs17331 (pathogenic for Osteogenesis imperfecta with normal sclerae, dominant form|Osteogenesis imperfecta type III) dbSNP  NCBI variation viewer
known disease mutation at this position, please check HGMD for details (HGMD ID CM960323)

known disease mutation at this position, please check HGMD for details (HGMD ID CM960323)
known disease mutation at this position, please check HGMD for details (HGMD ID CM960323)
regulatory features DNase1, Open Chromatin, DNase1 Hypersensitive Site
H3K18ac, Histone, Histone 3 Lysine 18 Acetylation
H3K27ac, Histone, Histone 3 Lysine 27 Acetylation
H3K27me3, Histone, Histone 3 Lysine 27 Tri-Methylation
H3K36me3, Histone, Histone 3 Lysine 36 Tri-Methylation
H3K9ac, Histone, Histone 3 Lysine 9 Acetylation
H4K5ac, Histone, Histone 4 Lysine 5 Acetylation
H4K91ac, Histone, Histone 4 Lysine 91 Acetylation
phyloP / phastCons
PhyloPPhastCons
(flanking)5.6961
5.6961
(flanking)3.1771
explain score(s) and/or inspect your position(s) in in UCSC Genome Browser
splice sites
effectgDNA positionscorewt detection sequence exon-intron border
Donor marginally increased11579wt: 0.9912 / mu: 0.9938 (marginal change - not scored)wt: CTAAAGGCGATGCTG
mu: CTAAAGGCGATGCTA
 AAAG|gcga
Donor gained115830.92mu: AGGCGATGCTAGTCC GCGA|tgct
distance from splice site 45
Kozak consensus sequence altered? N/A
conservation
protein level for non-synonymous changes
speciesmatchgeneaaalignment
Human      839EPGDAGAKGDAGPPGPAGPAGPPG
mutated  not conserved    839EPGDAGAKGDASPPGPAGPAGPP
Ptroglodytes  all identical  ENSPTRG00000009393  839EPGDAGAKGDAGPPGPAGPAGPP
Mmulatta  all identical  ENSMMUG00000001467  839EPGDAGAKGDAGPPGPAGPAGPP
Fcatus  no homologue    
Mmusculus  all identical  ENSMUSG00000001506  828EPGDTGVKGDAGPPGPAGPAGPP
Ggallus  no homologue    
Trubripes  all identical  ENSTRUG00000007520  831GPPGTTGPAGSS
Drerio  all identical  ENSDARG00000012405  823EPGDNGAKGDAGAPGPAGATGAP
Dmelanogaster  no homologue    
Celegans  no homologue    
Xtropicalis  all identical  ENSXETG00000003374  825EQGDAGAKGDAGPPGPAGPTGAP
protein features
start (aa)end (aa)featuredetails 
1791192REGIONTriple-helical region.lost
953954SITECleavage; by collagenase (By similarity).might get lost (downstream of altered splice site)
966968STRANDmight get lost (downstream of altered splice site)
975976CONFLICTLP -> PL (in Ref. 19; AAA52291).might get lost (downstream of altered splice site)
10811081CONFLICTV -> A (in Ref. 18; AAA51995).might get lost (downstream of altered splice site)
10931095MOTIFCell attachment site (Potential).might get lost (downstream of altered splice site)
11081108CARBOHYDO-linked (Gal...) (By similarity).might get lost (downstream of altered splice site)
11081108MOD_RES5-hydroxylysine (By similarity).might get lost (downstream of altered splice site)
11641164MOD_RES3-hydroxyproline (By similarity).might get lost (downstream of altered splice site)
11931218REGIONNonhelical region (C-terminal).might get lost (downstream of altered splice site)
12081208MOD_RESAllysine (By similarity).might get lost (downstream of altered splice site)
12181219SITECleavage; by procollagen C-endopeptidase.might get lost (downstream of altered splice site)
12191464PROPEPC-terminal propeptide. /FTId=PRO_0000005721.might get lost (downstream of altered splice site)
12291464DOMAINFibrillar collagen NC1.might get lost (downstream of altered splice site)
12591259DISULFIDInterchain (By similarity).might get lost (downstream of altered splice site)
12591259DISULFIDInterchain (By similarity).might get lost (downstream of altered splice site)
12651265DISULFIDInterchain (By similarity).might get lost (downstream of altered splice site)
12651265DISULFIDInterchain (By similarity).might get lost (downstream of altered splice site)
12821282DISULFIDInterchain (By similarity).might get lost (downstream of altered splice site)
12821282DISULFIDInterchain (By similarity).might get lost (downstream of altered splice site)
12911291DISULFIDInterchain (By similarity).might get lost (downstream of altered splice site)
12911291DISULFIDInterchain (By similarity).might get lost (downstream of altered splice site)
12991299DISULFIDBy similarity.might get lost (downstream of altered splice site)
13291329CONFLICTS -> T (in Ref. 25; AAB27856).might get lost (downstream of altered splice site)
13651365CARBOHYDN-linked (GlcNAc...).might get lost (downstream of altered splice site)
13701370DISULFIDBy similarity.might get lost (downstream of altered splice site)
14151415DISULFIDBy similarity.might get lost (downstream of altered splice site)
14621462DISULFIDBy similarity.might get lost (downstream of altered splice site)
length of protein normal
AA sequence altered yes
position of stopcodon in wt / mu CDS 4395 / 4395
position (AA) of stopcodon in wt / mu AA sequence 1465 / 1465
position of stopcodon in wt / mu cDNA 4514 / 4514
poly(A) signal N/A
conservation
nucleotide level for all changes - no scoring up to now
N/A
position of start ATG in wt / mu cDNA 120 / 120
chromosome 17
strand -1
last intron/exon boundary 4368
theoretical NMD boundary in CDS 4198
length of CDS 4395
coding sequence (CDS) position 2515
cDNA position
(for ins/del: last normal base / first normal base)
2634
gDNA position
(for ins/del: last normal base / first normal base)
11588
chromosomal position
(for ins/del: last normal base / first normal base)
48267406
original gDNA sequence snippet CTGGTGCTAAAGGCGATGCTGGTCCCCCTGGCCCTGCCGGA
altered gDNA sequence snippet CTGGTGCTAAAGGCGATGCTAGTCCCCCTGGCCCTGCCGGA
original cDNA sequence snippet CTGGTGCTAAAGGCGATGCTGGTCCCCCTGGCCCTGCCGGA
altered cDNA sequence snippet CTGGTGCTAAAGGCGATGCTAGTCCCCCTGGCCCTGCCGGA
wildtype AA sequence MFSFVDLRLL LLLAATALLT HGQEEGQVEG QDEDIPPITC VQNGLRYHDR DVWKPEPCRI
CVCDNGKVLC DDVICDETKN CPGAEVPEGE CCPVCPDGSE SPTDQETTGV EGPKGDTGPR
GPRGPAGPPG RDGIPGQPGL PGPPGPPGPP GPPGLGGNFA PQLSYGYDEK STGGISVPGP
MGPSGPRGLP GPPGAPGPQG FQGPPGEPGE PGASGPMGPR GPPGPPGKNG DDGEAGKPGR
PGERGPPGPQ GARGLPGTAG LPGMKGHRGF SGLDGAKGDA GPAGPKGEPG SPGENGAPGQ
MGPRGLPGER GRPGAPGPAG ARGNDGATGA AGPPGPTGPA GPPGFPGAVG AKGEAGPQGP
RGSEGPQGVR GEPGPPGPAG AAGPAGNPGA DGQPGAKGAN GAPGIAGAPG FPGARGPSGP
QGPGGPPGPK GNSGEPGAPG SKGDTGAKGE PGPVGVQGPP GPAGEEGKRG ARGEPGPTGL
PGPPGERGGP GSRGFPGADG VAGPKGPAGE RGSPGPAGPK GSPGEAGRPG EAGLPGAKGL
TGSPGSPGPD GKTGPPGPAG QDGRPGPPGP PGARGQAGVM GFPGPKGAAG EPGKAGERGV
PGPPGAVGPA GKDGEAGAQG PPGPAGPAGE RGEQGPAGSP GFQGLPGPAG PPGEAGKPGE
QGVPGDLGAP GPSGARGERG FPGERGVQGP PGPAGPRGAN GAPGNDGAKG DAGAPGAPGS
QGAPGLQGMP GERGAAGLPG PKGDRGDAGP KGADGSPGKD GVRGLTGPIG PPGPAGAPGD
KGESGPSGPA GPTGARGAPG DRGEPGPPGP AGFAGPPGAD GQPGAKGEPG DAGAKGDAGP
PGPAGPAGPP GPIGNVGAPG AKGARGSAGP PGATGFPGAA GRVGPPGPSG NAGPPGPPGP
AGKEGGKGPR GETGPAGRPG EVGPPGPPGP AGEKGSPGAD GPAGAPGTPG PQGIAGQRGV
VGLPGQRGER GFPGLPGPSG EPGKQGPSGA SGERGPPGPM GPPGLAGPPG ESGREGAPGA
EGSPGRDGSP GAKGDRGETG PAGPPGAPGA PGAPGPVGPA GKSGDRGETG PAGPTGPVGP
VGARGPAGPQ GPRGDKGETG EQGDRGIKGH RGFSGLQGPP GPPGSPGEQG PSGASGPAGP
RGPPGSAGAP GKDGLNGLPG PIGPPGPRGR TGDAGPVGPP GPPGPPGPPG PPSAGFDFSF
LPQPPQEKAH DGGRYYRADD ANVVRDRDLE VDTTLKSLSQ QIENIRSPEG SRKNPARTCR
DLKMCHSDWK SGEYWIDPNQ GCNLDAIKVF CNMETGETCV YPTQPSVAQK NWYISKNPKD
KRHVWFGESM TDGFQFEYGG QGSDPADVAI QLTFLRLMST EASQNITYHC KNSVAYMDQQ
TGNLKKALLL QGSNEIEIRA EGNSRFTYSV TVDGCTSHTG AWGKTVIEYK TTKTSRLPII
DVAPLDVGAP DQEFGFDVGP VCFL*
mutated AA sequence MFSFVDLRLL LLLAATALLT HGQEEGQVEG QDEDIPPITC VQNGLRYHDR DVWKPEPCRI
CVCDNGKVLC DDVICDETKN CPGAEVPEGE CCPVCPDGSE SPTDQETTGV EGPKGDTGPR
GPRGPAGPPG RDGIPGQPGL PGPPGPPGPP GPPGLGGNFA PQLSYGYDEK STGGISVPGP
MGPSGPRGLP GPPGAPGPQG FQGPPGEPGE PGASGPMGPR GPPGPPGKNG DDGEAGKPGR
PGERGPPGPQ GARGLPGTAG LPGMKGHRGF SGLDGAKGDA GPAGPKGEPG SPGENGAPGQ
MGPRGLPGER GRPGAPGPAG ARGNDGATGA AGPPGPTGPA GPPGFPGAVG AKGEAGPQGP
RGSEGPQGVR GEPGPPGPAG AAGPAGNPGA DGQPGAKGAN GAPGIAGAPG FPGARGPSGP
QGPGGPPGPK GNSGEPGAPG SKGDTGAKGE PGPVGVQGPP GPAGEEGKRG ARGEPGPTGL
PGPPGERGGP GSRGFPGADG VAGPKGPAGE RGSPGPAGPK GSPGEAGRPG EAGLPGAKGL
TGSPGSPGPD GKTGPPGPAG QDGRPGPPGP PGARGQAGVM GFPGPKGAAG EPGKAGERGV
PGPPGAVGPA GKDGEAGAQG PPGPAGPAGE RGEQGPAGSP GFQGLPGPAG PPGEAGKPGE
QGVPGDLGAP GPSGARGERG FPGERGVQGP PGPAGPRGAN GAPGNDGAKG DAGAPGAPGS
QGAPGLQGMP GERGAAGLPG PKGDRGDAGP KGADGSPGKD GVRGLTGPIG PPGPAGAPGD
KGESGPSGPA GPTGARGAPG DRGEPGPPGP AGFAGPPGAD GQPGAKGEPG DAGAKGDASP
PGPAGPAGPP GPIGNVGAPG AKGARGSAGP PGATGFPGAA GRVGPPGPSG NAGPPGPPGP
AGKEGGKGPR GETGPAGRPG EVGPPGPPGP AGEKGSPGAD GPAGAPGTPG PQGIAGQRGV
VGLPGQRGER GFPGLPGPSG EPGKQGPSGA SGERGPPGPM GPPGLAGPPG ESGREGAPGA
EGSPGRDGSP GAKGDRGETG PAGPPGAPGA PGAPGPVGPA GKSGDRGETG PAGPTGPVGP
VGARGPAGPQ GPRGDKGETG EQGDRGIKGH RGFSGLQGPP GPPGSPGEQG PSGASGPAGP
RGPPGSAGAP GKDGLNGLPG PIGPPGPRGR TGDAGPVGPP GPPGPPGPPG PPSAGFDFSF
LPQPPQEKAH DGGRYYRADD ANVVRDRDLE VDTTLKSLSQ QIENIRSPEG SRKNPARTCR
DLKMCHSDWK SGEYWIDPNQ GCNLDAIKVF CNMETGETCV YPTQPSVAQK NWYISKNPKD
KRHVWFGESM TDGFQFEYGG QGSDPADVAI QLTFLRLMST EASQNITYHC KNSVAYMDQQ
TGNLKKALLL QGSNEIEIRA EGNSRFTYSV TVDGCTSHTG AWGKTVIEYK TTKTSRLPII
DVAPLDVGAP DQEFGFDVGP VCFL*
speed 1.33 s
All positions are in basepairs (bp) if not explicitly stated differently.
AA/aa: amino acid; CDS: coding sequence; mu: mutated; NMD: nonsense-mediated mRNA decay; nt: nucleotide; wt: wildtype; TGP: 1000 Genomes Project
back to results table

Problems