Amino acid dipepetide frequency for Streptococcus satellite phage Javan442

Apart from single amino acid frequencies one can also calculate so called amino acid dipeptide frequency.

There is 400 possibilites (441 if we consider Xaa as additional 21st amino acid representing all non-standard or unknown amino acids). Thus, if the dipeptides would be present randomly in proteins each amino acid dipeptide should be present with 0.25% frequency. As it is not the case in the nature, for better visablity all more than expected dipepetides are marked by red, and those which are underrepresented are marked by blue in the table.

All values are presented as per milles (‰), therefore need to be multiplied by 10-3.

For more information see sequence space article on Wikipedia.

AlaCysAspGluPheGlyHisIleLysLeuMetAsnProGlnArgSerThrValTrpTyrXaa
Ala
0.0AlaAla: 0.0 ± 0.0
0.499AlaCys: 0.499 ± 0.456
2.991AlaAsp: 2.991 ± 0.959
3.49AlaGlu: 3.49 ± 1.034
3.49AlaPhe: 3.49 ± 0.857
1.496AlaGly: 1.496 ± 1.084
0.0AlaHis: 0.0 ± 0.0
5.982AlaIle: 5.982 ± 1.678
4.487AlaLys: 4.487 ± 1.545
6.481AlaLeu: 6.481 ± 1.61
2.991AlaMet: 2.991 ± 1.276
0.997AlaAsn: 0.997 ± 0.92
1.994AlaPro: 1.994 ± 0.842
1.994AlaGln: 1.994 ± 0.971
2.991AlaArg: 2.991 ± 1.316
1.994AlaSer: 1.994 ± 0.886
2.991AlaThr: 2.991 ± 1.377
1.994AlaVal: 1.994 ± 0.793
0.0AlaTrp: 0.0 ± 0.0
1.496AlaTyr: 1.496 ± 0.984
0.0AlaXaa: 0.0 ± 0.0
Cys
0.0CysAla: 0.0 ± 0.0
0.0CysCys: 0.0 ± 0.0
0.499CysAsp: 0.499 ± 0.467
0.997CysGlu: 0.997 ± 0.552
0.0CysPhe: 0.0 ± 0.0
1.994CysGly: 1.994 ± 1.037
0.0CysHis: 0.0 ± 0.0
0.499CysIle: 0.499 ± 0.46
0.997CysLys: 0.997 ± 0.787
1.496CysLeu: 1.496 ± 1.019
0.0CysMet: 0.0 ± 0.0
0.0CysAsn: 0.0 ± 0.0
0.0CysPro: 0.0 ± 0.0
0.0CysGln: 0.0 ± 0.0
0.0CysArg: 0.0 ± 0.0
0.499CysSer: 0.499 ± 0.427
0.499CysThr: 0.499 ± 0.456
0.997CysVal: 0.997 ± 0.912
0.0CysTrp: 0.0 ± 0.0
1.496CysTyr: 1.496 ± 0.865
0.0CysXaa: 0.0 ± 0.0
Asp
1.994AspAla: 1.994 ± 0.884
0.997AspCys: 0.997 ± 0.912
3.49AspAsp: 3.49 ± 1.334
3.988AspGlu: 3.988 ± 1.399
3.49AspPhe: 3.49 ± 0.947
1.994AspGly: 1.994 ± 0.918
0.0AspHis: 0.0 ± 0.0
6.979AspIle: 6.979 ± 1.807
6.979AspLys: 6.979 ± 2.056
5.982AspLeu: 5.982 ± 2.114
2.493AspMet: 2.493 ± 0.915
4.487AspAsn: 4.487 ± 1.472
0.997AspPro: 0.997 ± 0.541
1.496AspGln: 1.496 ± 0.774
2.991AspArg: 2.991 ± 0.982
1.994AspSer: 1.994 ± 0.883
1.994AspThr: 1.994 ± 1.044
0.997AspVal: 0.997 ± 0.68
0.997AspTrp: 0.997 ± 0.642
4.985AspTyr: 4.985 ± 1.935
0.0AspXaa: 0.0 ± 0.0
Glu
6.481GluAla: 6.481 ± 1.852
0.997GluCys: 0.997 ± 0.674
6.481GluAsp: 6.481 ± 1.956
6.481GluGlu: 6.481 ± 1.715
2.991GluPhe: 2.991 ± 1.142
1.496GluGly: 1.496 ± 0.783
0.997GluHis: 0.997 ± 0.628
2.991GluIle: 2.991 ± 1.034
8.475GluLys: 8.475 ± 1.34
12.463GluLeu: 12.463 ± 1.933
1.496GluMet: 1.496 ± 0.904
3.49GluAsn: 3.49 ± 1.737
1.994GluPro: 1.994 ± 0.809
2.991GluGln: 2.991 ± 1.357
4.487GluArg: 4.487 ± 0.874
1.496GluSer: 1.496 ± 0.712
1.496GluThr: 1.496 ± 0.865
3.49GluVal: 3.49 ± 1.378
0.997GluTrp: 0.997 ± 0.655
4.487GluTyr: 4.487 ± 1.51
0.0GluXaa: 0.0 ± 0.0
Phe
1.994PheAla: 1.994 ± 0.919
0.499PheCys: 0.499 ± 0.456
1.994PheAsp: 1.994 ± 0.899
4.985PheGlu: 4.985 ± 1.76
1.994PhePhe: 1.994 ± 1.168
1.994PheGly: 1.994 ± 1.241
1.496PheHis: 1.496 ± 1.08
3.49PheIle: 3.49 ± 1.025
4.487PheLys: 4.487 ± 1.901
1.994PheLeu: 1.994 ± 0.84
0.997PheMet: 0.997 ± 0.78
0.499PheAsn: 0.499 ± 0.537
1.496PhePro: 1.496 ± 0.935
0.499PheGln: 0.499 ± 0.547
2.493PheArg: 2.493 ± 1.168
1.994PheSer: 1.994 ± 1.012
1.994PheThr: 1.994 ± 0.932
0.997PheVal: 0.997 ± 1.073
0.0PheTrp: 0.0 ± 0.0
0.997PheTyr: 0.997 ± 0.509
0.0PheXaa: 0.0 ± 0.0
Gly
1.994GlyAla: 1.994 ± 0.933
1.496GlyCys: 1.496 ± 0.791
1.994GlyAsp: 1.994 ± 1.094
0.499GlyGlu: 0.499 ± 0.417
1.496GlyPhe: 1.496 ± 0.677
2.991GlyGly: 2.991 ± 1.02
0.997GlyHis: 0.997 ± 0.74
4.985GlyIle: 4.985 ± 0.982
4.487GlyLys: 4.487 ± 1.219
3.49GlyLeu: 3.49 ± 1.071
0.997GlyMet: 0.997 ± 0.645
2.991GlyAsn: 2.991 ± 1.027
0.0GlyPro: 0.0 ± 0.0
1.496GlyGln: 1.496 ± 0.839
2.991GlyArg: 2.991 ± 1.109
2.991GlySer: 2.991 ± 1.21
3.49GlyThr: 3.49 ± 1.16
3.988GlyVal: 3.988 ± 1.091
0.0GlyTrp: 0.0 ± 0.0
4.487GlyTyr: 4.487 ± 1.329
0.0GlyXaa: 0.0 ± 0.0
His
1.994HisAla: 1.994 ± 1.22
0.0HisCys: 0.0 ± 0.0
0.0HisAsp: 0.0 ± 0.0
2.493HisGlu: 2.493 ± 1.074
0.499HisPhe: 0.499 ± 0.503
0.0HisGly: 0.0 ± 0.0
0.499HisHis: 0.499 ± 0.547
0.499HisIle: 0.499 ± 0.574
0.997HisLys: 0.997 ± 0.751
1.496HisLeu: 1.496 ± 0.766
0.0HisMet: 0.0 ± 0.0
0.0HisAsn: 0.0 ± 0.0
0.0HisPro: 0.0 ± 0.0
0.499HisGln: 0.499 ± 0.431
0.997HisArg: 0.997 ± 0.635
0.997HisSer: 0.997 ± 0.679
1.496HisThr: 1.496 ± 0.894
0.997HisVal: 0.997 ± 0.621
0.0HisTrp: 0.0 ± 0.0
1.994HisTyr: 1.994 ± 0.95
0.0HisXaa: 0.0 ± 0.0
Ile
3.988IleAla: 3.988 ± 1.137
2.493IleCys: 2.493 ± 1.054
6.481IleAsp: 6.481 ± 1.868
2.493IleGlu: 2.493 ± 1.024
0.499IlePhe: 0.499 ± 0.576
4.487IleGly: 4.487 ± 1.597
1.496IleHis: 1.496 ± 0.892
4.487IleIle: 4.487 ± 1.176
9.472IleLys: 9.472 ± 2.508
7.478IleLeu: 7.478 ± 2.128
1.994IleMet: 1.994 ± 0.994
5.982IleAsn: 5.982 ± 1.374
3.49IlePro: 3.49 ± 1.033
0.997IleGln: 0.997 ± 0.621
1.994IleArg: 1.994 ± 1.334
5.484IleSer: 5.484 ± 1.887
6.481IleThr: 6.481 ± 1.893
2.991IleVal: 2.991 ± 1.125
0.0IleTrp: 0.0 ± 0.0
2.991IleTyr: 2.991 ± 1.041
0.0IleXaa: 0.0 ± 0.0
Lys
6.481LysAla: 6.481 ± 1.854
0.0LysCys: 0.0 ± 0.0
5.484LysAsp: 5.484 ± 1.484
8.475LysGlu: 8.475 ± 1.988
0.997LysPhe: 0.997 ± 0.715
3.49LysGly: 3.49 ± 1.142
2.493LysHis: 2.493 ± 1.212
6.481LysIle: 6.481 ± 1.628
11.964LysLys: 11.964 ± 2.656
11.964LysLeu: 11.964 ± 2.529
3.49LysMet: 3.49 ± 1.271
6.481LysAsn: 6.481 ± 1.541
2.493LysPro: 2.493 ± 0.744
5.484LysGln: 5.484 ± 1.629
8.973LysArg: 8.973 ± 2.012
4.487LysSer: 4.487 ± 1.136
4.487LysThr: 4.487 ± 1.514
4.487LysVal: 4.487 ± 0.996
0.499LysTrp: 0.499 ± 0.431
5.484LysTyr: 5.484 ± 1.678
0.0LysXaa: 0.0 ± 0.0
Leu
6.481LeuAla: 6.481 ± 1.502
0.0LeuCys: 0.0 ± 0.0
6.979LeuAsp: 6.979 ± 1.323
10.469LeuGlu: 10.469 ± 2.177
2.991LeuPhe: 2.991 ± 1.482
8.475LeuGly: 8.475 ± 1.811
1.496LeuHis: 1.496 ± 0.779
6.481LeuIle: 6.481 ± 1.899
9.97LeuLys: 9.97 ± 2.304
4.985LeuLeu: 4.985 ± 2.071
5.484LeuMet: 5.484 ± 1.212
7.478LeuAsn: 7.478 ± 1.44
2.493LeuPro: 2.493 ± 1.214
3.49LeuGln: 3.49 ± 1.198
4.487LeuArg: 4.487 ± 1.351
4.985LeuSer: 4.985 ± 0.904
6.481LeuThr: 6.481 ± 1.42
5.484LeuVal: 5.484 ± 1.599
2.493LeuTrp: 2.493 ± 0.951
1.994LeuTyr: 1.994 ± 0.758
0.0LeuXaa: 0.0 ± 0.0
Met
2.991MetAla: 2.991 ± 1.198
0.0MetCys: 0.0 ± 0.0
1.994MetAsp: 1.994 ± 1.155
2.991MetGlu: 2.991 ± 1.03
0.0MetPhe: 0.0 ± 0.0
0.499MetGly: 0.499 ± 0.417
0.0MetHis: 0.0 ± 0.0
0.997MetIle: 0.997 ± 1.071
4.487MetLys: 4.487 ± 1.717
1.994MetLeu: 1.994 ± 0.965
0.499MetMet: 0.499 ± 0.617
1.994MetAsn: 1.994 ± 0.851
0.997MetPro: 0.997 ± 0.628
1.496MetGln: 1.496 ± 0.952
0.997MetArg: 0.997 ± 0.654
1.994MetSer: 1.994 ± 0.858
2.493MetThr: 2.493 ± 1.058
0.997MetVal: 0.997 ± 0.523
0.0MetTrp: 0.0 ± 0.0
0.499MetTyr: 0.499 ± 0.507
0.0MetXaa: 0.0 ± 0.0
Asn
1.994AsnAla: 1.994 ± 1.105
0.0AsnCys: 0.0 ± 0.0
2.493AsnAsp: 2.493 ± 1.393
3.49AsnGlu: 3.49 ± 1.244
2.493AsnPhe: 2.493 ± 1.18
2.991AsnGly: 2.991 ± 1.085
0.0AsnHis: 0.0 ± 0.0
4.985AsnIle: 4.985 ± 1.041
6.481AsnLys: 6.481 ± 1.534
3.988AsnLeu: 3.988 ± 1.355
1.496AsnMet: 1.496 ± 0.841
3.988AsnAsn: 3.988 ± 1.447
1.994AsnPro: 1.994 ± 0.919
2.991AsnGln: 2.991 ± 1.092
1.994AsnArg: 1.994 ± 0.912
2.991AsnSer: 2.991 ± 1.03
2.493AsnThr: 2.493 ± 1.244
1.496AsnVal: 1.496 ± 0.99
1.496AsnTrp: 1.496 ± 0.812
2.493AsnTyr: 2.493 ± 1.293
0.0AsnXaa: 0.0 ± 0.0
Pro
0.997ProAla: 0.997 ± 0.706
0.499ProCys: 0.499 ± 0.431
0.997ProAsp: 0.997 ± 0.547
2.493ProGlu: 2.493 ± 0.981
2.493ProPhe: 2.493 ± 1.123
0.0ProGly: 0.0 ± 0.0
0.499ProHis: 0.499 ± 0.503
1.496ProIle: 1.496 ± 0.694
4.487ProLys: 4.487 ± 1.165
0.997ProLeu: 0.997 ± 0.651
0.499ProMet: 0.499 ± 0.456
0.499ProAsn: 0.499 ± 0.547
0.997ProPro: 0.997 ± 0.646
1.496ProGln: 1.496 ± 0.848
2.493ProArg: 2.493 ± 0.875
0.499ProSer: 0.499 ± 0.456
1.496ProThr: 1.496 ± 0.997
0.499ProVal: 0.499 ± 0.511
0.0ProTrp: 0.0 ± 0.0
0.499ProTyr: 0.499 ± 0.547
0.0ProXaa: 0.0 ± 0.0
Gln
2.991GlnAla: 2.991 ± 0.93
0.499GlnCys: 0.499 ± 0.617
2.493GlnAsp: 2.493 ± 1.262
3.988GlnGlu: 3.988 ± 1.115
0.997GlnPhe: 0.997 ± 0.654
1.994GlnGly: 1.994 ± 1.137
0.0GlnHis: 0.0 ± 0.0
3.988GlnIle: 3.988 ± 1.279
2.991GlnLys: 2.991 ± 1.308
7.478GlnLeu: 7.478 ± 2.318
1.496GlnMet: 1.496 ± 0.614
1.496GlnAsn: 1.496 ± 0.764
0.0GlnPro: 0.0 ± 0.0
3.49GlnGln: 3.49 ± 1.307
0.997GlnArg: 0.997 ± 0.721
3.988GlnSer: 3.988 ± 1.529
0.997GlnThr: 0.997 ± 0.599
1.994GlnVal: 1.994 ± 0.796
0.0GlnTrp: 0.0 ± 0.0
3.49GlnTyr: 3.49 ± 1.213
0.0GlnXaa: 0.0 ± 0.0
Arg
0.997ArgAla: 0.997 ± 0.626
0.499ArgCys: 0.499 ± 0.456
2.991ArgAsp: 2.991 ± 1.487
4.985ArgGlu: 4.985 ± 1.63
0.997ArgPhe: 0.997 ± 0.609
3.49ArgGly: 3.49 ± 1.051
0.997ArgHis: 0.997 ± 0.657
2.493ArgIle: 2.493 ± 1.039
3.988ArgLys: 3.988 ± 1.236
8.475ArgLeu: 8.475 ± 1.884
0.0ArgMet: 0.0 ± 0.0
1.496ArgAsn: 1.496 ± 0.806
1.496ArgPro: 1.496 ± 0.62
3.49ArgGln: 3.49 ± 1.316
2.493ArgArg: 2.493 ± 0.964
1.994ArgSer: 1.994 ± 0.919
4.487ArgThr: 4.487 ± 2.1
5.982ArgVal: 5.982 ± 1.018
0.997ArgTrp: 0.997 ± 0.509
2.493ArgTyr: 2.493 ± 1.102
0.0ArgXaa: 0.0 ± 0.0
Ser
1.994SerAla: 1.994 ± 1.252
0.499SerCys: 0.499 ± 0.503
3.49SerAsp: 3.49 ± 1.072
3.988SerGlu: 3.988 ± 1.422
3.49SerPhe: 3.49 ± 1.564
2.493SerGly: 2.493 ± 1.086
0.499SerHis: 0.499 ± 0.503
7.478SerIle: 7.478 ± 1.487
2.991SerLys: 2.991 ± 1.273
5.982SerLeu: 5.982 ± 1.336
0.499SerMet: 0.499 ± 0.517
1.994SerAsn: 1.994 ± 0.983
0.0SerPro: 0.0 ± 0.0
3.49SerGln: 3.49 ± 0.859
1.496SerArg: 1.496 ± 0.716
1.496SerSer: 1.496 ± 0.912
3.988SerThr: 3.988 ± 1.153
3.988SerVal: 3.988 ± 1.473
0.0SerTrp: 0.0 ± 0.0
3.49SerTyr: 3.49 ± 1.308
0.0SerXaa: 0.0 ± 0.0
Thr
0.997ThrAla: 0.997 ± 0.694
0.0ThrCys: 0.0 ± 0.0
3.49ThrAsp: 3.49 ± 1.135
2.991ThrGlu: 2.991 ± 1.039
3.49ThrPhe: 3.49 ± 1.106
3.49ThrGly: 3.49 ± 1.249
1.496ThrHis: 1.496 ± 0.77
5.982ThrIle: 5.982 ± 1.444
4.985ThrLys: 4.985 ± 1.541
7.976ThrLeu: 7.976 ± 1.522
0.499ThrMet: 0.499 ± 0.536
1.994ThrAsn: 1.994 ± 1.037
2.493ThrPro: 2.493 ± 0.872
2.991ThrGln: 2.991 ± 1.232
3.988ThrArg: 3.988 ± 1.385
1.496ThrSer: 1.496 ± 0.8
1.994ThrThr: 1.994 ± 0.743
3.988ThrVal: 3.988 ± 1.19
0.499ThrTrp: 0.499 ± 0.472
1.496ThrTyr: 1.496 ± 0.721
0.0ThrXaa: 0.0 ± 0.0
Val
2.991ValAla: 2.991 ± 0.988
0.499ValCys: 0.499 ± 0.431
3.49ValAsp: 3.49 ± 0.925
1.496ValGlu: 1.496 ± 0.893
1.994ValPhe: 1.994 ± 1.015
2.493ValGly: 2.493 ± 1.049
1.496ValHis: 1.496 ± 0.739
2.991ValIle: 2.991 ± 1.384
3.988ValLys: 3.988 ± 1.326
3.49ValLeu: 3.49 ± 1.266
0.499ValMet: 0.499 ± 0.576
3.49ValAsn: 3.49 ± 1.724
0.499ValPro: 0.499 ± 0.431
2.991ValGln: 2.991 ± 1.106
2.991ValArg: 2.991 ± 1.215
5.982ValSer: 5.982 ± 1.519
4.487ValThr: 4.487 ± 1.289
1.994ValVal: 1.994 ± 0.951
0.0ValTrp: 0.0 ± 0.0
1.496ValTyr: 1.496 ± 0.912
0.0ValXaa: 0.0 ± 0.0
Trp
0.0TrpAla: 0.0 ± 0.0
0.0TrpCys: 0.0 ± 0.0
0.997TrpAsp: 0.997 ± 0.577
2.493TrpGlu: 2.493 ± 1.311
0.0TrpPhe: 0.0 ± 0.0
0.997TrpGly: 0.997 ± 0.547
0.499TrpHis: 0.499 ± 0.431
0.0TrpIle: 0.0 ± 0.0
0.0TrpLys: 0.0 ± 0.0
0.997TrpLeu: 0.997 ± 0.67
0.0TrpMet: 0.0 ± 0.0
0.0TrpAsn: 0.0 ± 0.0
0.0TrpPro: 0.0 ± 0.0
0.499TrpGln: 0.499 ± 0.431
0.499TrpArg: 0.499 ± 0.467
0.997TrpSer: 0.997 ± 0.662
0.0TrpThr: 0.0 ± 0.0
0.0TrpVal: 0.0 ± 0.0
0.0TrpTrp: 0.0 ± 0.0
0.997TrpTyr: 0.997 ± 0.855
0.0TrpXaa: 0.0 ± 0.0
Tyr
0.997TyrAla: 0.997 ± 0.753
0.499TyrCys: 0.499 ± 0.428
0.499TyrAsp: 0.499 ± 0.576
2.991TyrGlu: 2.991 ± 0.885
2.493TyrPhe: 2.493 ± 1.224
0.997TyrGly: 0.997 ± 0.684
0.499TyrHis: 0.499 ± 0.537
2.991TyrIle: 2.991 ± 0.913
7.478TyrLys: 7.478 ± 1.721
3.988TyrLeu: 3.988 ± 1.147
1.994TyrMet: 1.994 ± 1.825
2.991TyrAsn: 2.991 ± 0.9
0.499TyrPro: 0.499 ± 0.617
3.49TyrGln: 3.49 ± 0.936
4.487TyrArg: 4.487 ± 1.264
4.985TyrSer: 4.985 ± 1.198
2.493TyrThr: 2.493 ± 1.243
1.994TyrVal: 1.994 ± 1.114
0.997TyrTrp: 0.997 ± 0.577
2.991TyrTyr: 2.991 ± 1.307
0.0TyrXaa: 0.0 ± 0.0
Xaa
0.0XaaAla: 0.0 ± 0.0
0.0XaaCys: 0.0 ± 0.0
0.0XaaAsp: 0.0 ± 0.0
0.0XaaGlu: 0.0 ± 0.0
0.0XaaPhe: 0.0 ± 0.0
0.0XaaGly: 0.0 ± 0.0
0.0XaaHis: 0.0 ± 0.0
0.0XaaIle: 0.0 ± 0.0
0.0XaaLys: 0.0 ± 0.0
0.0XaaLeu: 0.0 ± 0.0
0.0XaaMet: 0.0 ± 0.0
0.0XaaAsn: 0.0 ± 0.0
0.0XaaPro: 0.0 ± 0.0
0.0XaaGln: 0.0 ± 0.0
0.0XaaArg: 0.0 ± 0.0
0.0XaaSer: 0.0 ± 0.0
0.0XaaThr: 0.0 ± 0.0
0.0XaaVal: 0.0 ± 0.0
0.0XaaTrp: 0.0 ± 0.0
0.0XaaTyr: 0.0 ± 0.0
0.0XaaXaa: 0.0 ± 0.0
Statistics based on 19 proteins (2007 amino acids)

Note: The error has been estimated with the bootstraping (x100) at the protein level

Above dipeptide statistics (among other stats for this proteome) you can download from this CSV file
See this proteome in: uniprot_link
Proteome-pI is available under Creative Commons Attribution-NoDerivs license, for more details see here

Reference: Kozlowski LP. Proteome-pI 2.0: Proteome Isoelectric Point Database Update. Nucleic Acids Res. 2021, doi: 10.1093/nar/gkab944 Contact: Lukasz P. Kozlowski