Wednesday, September 16, 2020

Study: Southwestern Finnish samples from the late Viking Age to the 1200s

A new study has landed on the earth, including 6 Finnish female samples and 2 Finnish male samples. According results Finnish samples look very similar to modern Southwestern Finns, Ukrainians and Belarusians. Two male included belonging to the haplogroup N, but results are not very informative, because they are at same level with ancient Kola Peninsula samples (Bolshoy Oleni Ostrov samples).  There is also a close-up PCA plot showing close relation to medieval Estonian samples from Saag et al. 2019 and two Swedish Viking Age samples from Sigtuna, namely kal006 and stg020. The study gives different result than the original study (Maja Krzewinska et al. 2018: Genomic and Strontium Isotope Variation Reveal Immigration Patterns in a Viking Age Town) for stg020, because stg020 from the Church 1 burial is according to the original study an outlier from East Europe (Ukraine?), anyway not a Finnish migrant. Who knows what is the truth...  No genome data publicly available, as usual in case of Finnish studies.  

https://eaa.klinkhamergroup.com/eaa2...q51ZoW7lGj99c0


Edit 20.4.21 11:00

The study has disappeared.  I would be thankful if readers could help me to find it again.

Tuesday, August 18, 2020

Popular Vahaduo Oracle for Finns

I made a small oracle experience using ancient dna.  It is a simple exercize giving possibility to get personal admixture results using the Vahaduo online service.  You need to have your own Eurogenes K13 results from Gedmatch and place it and following source data to the Vahaduo service.  Just replace original source data by  following lines:

Estonian_IA,30.75,58.15,3.52,1.28,0.03,0.47,1.42,0.00,1.98,0.49,0.35,0.24,1.31
Levanluhta_IA,22.40,45.36,1.2,0.01,0.00,0.07,3.07,0.58,23.35,3.30,0.00,0.00,0.62
VolgaForestSteppe_CWC,35.78,37.16,0.02,18.65,0.00,0.00,3.55,0.00,1.51,2.51,0.00,0.00,0.82
Scandinavian_IA,54.70,30.08,6.46,2.79,0.00,1.71,1.35,0.00,0.41,1.71,0.00,0.04,0.77

Then place your K13 result as a target and calculate distant, single or multiple results dependnig on your own sample number.  Here is a multiple sample result:

 

Or you can use same data with your own oracle calculator instead of Vahaduo.

edit 20.8.2020 20:15

 

Sigtuna samples (Viking Age Swedish) without Volga Forest Steppe give better distances than Scandinavian Iron Age.  

Estonian_IA,30.75,58.15,3.52,1.28,0.03,0.47,1.42,0.00,1.98,0.49,0.35,0.24,1.31
Levanluhta_IA,22.40,45.36,1.2,0.01,0.00,0.07,3.07,0.58,23.35,3.30,0.00,0.00,0.62
Sigtuna,47.36,33.02,8.05,5.04,0.09,0.45,1.01,0.85,0.80,0.82,0.74,1.44,0.34



 

 

 

 

Friday, August 14, 2020

I-L258 brought Proto-Germanic language to Finland?

It is a well-known idea that Baltic-Finnic languages, including Finnish and Estonian, and Saami adopted hundreds Proto-Germanic loan words, but it is poorly known where it happened and even more poorly how it happened, because Germanic speakers lived in Scandinavia, Central and West Europe and Baltic Finnic and Saami speakers lived in the other side of the Baltic Sea.  This is all proven by historical sources and research of prehistory.   Even though many geneticists willingly search support from linguistic theories and studies for their own conclusion they have not solved this question.   It is also a no-brainer that conclusions made between genes and languages can't hold causation, thus proving genetic divergence doesn't call for linguistic parity and it is question about correlation and how exclusive it is.

So let's look where the loan word adaptation would have happened based on linguistic theories.  The most common idea for the Germanic urheimat is in the area area from North Germany to Southern Scandinavia.

https://en.wikipedia.org/wiki/Proto-Germanic_language

I think that this is the best estimation, although I have seen someones placing it to different places from present-day Russia to South Germany.   Reasons for these different opinions come probably from different datings of loan word and stratums in parallel languages and how methods in comparative linguistics treat these loan words.  In other words we have time, place and evolution of languages and all this depends on how the evolution goes on in time and place.  So, in some area the language was more diverged from the the  "proto-language", in some other places the language was still more in "pre-proto-language" stage.  This is normal linguistic evolution which happened and happens all the time and everywhere.  I can't imagine that some language was complete in some place at some particular time 


Let's then look the time.  The same Wikipedia article says that Proto-Germanic was spoken after 500BC and from 200AD, according to linguistics, its spoken successor is named Proto-Norse in the area we are now interested in.  In other areas the successors were something else. 


Then let's look to Baltic-Finnic languages.  The Wikipedia says that the diversification began in Southern Estonia 150AD, based on loan word evidences.  Some estimates give earlier dates, from 1000 - 600 BC.  Most of Baltic-Finnic branches are rather young, only 1000 years old.  But it is now enough to know that Baltic-Finnic languages were spoken in certain area from 500BC to 150AD, simultaneously with the Proto-Germanic language and the Wikipedia quote is:


" There is now wide agreement that Proto-Finnic was probably spoken at the coasts of the Gulf of Finland."


On the map we see a detached Karelian group, but it was born late in a historical era.


https://en.m.wikipedia.org/wiki/Finnic_languages#:~:text=The%20Finnic%20(Fennic)%20or%20Balto,Finnic%20languages%20have%20been%20recognized


The local history and birth of Saami language in Baltic area as explained Wikipedia:


"Proto-Samic language is believed to have formed in the vicinity of the Gulf of Finland between 1000 BC to 700 AD, deriving from a common Proto-Sami-Finnic language (M. Korhonen 1981)."

 

https://en.m.wikipedia.org/wiki/S%C3%A1mi_languages

 

It is important to know that also the Saami was spoken in a nearby area of the Proto-Germanic language and simultaneously with it. 


Let's then do the same thing geneticists use to do.  We can ask which widely known and proven genetic results would get the best chance for the Proto-Germanic loan word stratum found  in Baltic-Finnic and Saami languages and in donating and receiving languages fitting with the place and dating.  I made the following chart showing datings and places of the Scandinavian I1 haplogroup L22.  All allele mutation data is gathered from FamilytreeDna and presents obviously BIG-Y results.  All datings are from Yfull.

edit.  correction made regarding statistics.  For instance the number in column I-Z74 includes, inter alia, tested L813 samples, but not the parallel downstream clade I-CTS2208.  

edit.  some numbers corrected. 

edit. added Italian L22.


Monday, August 3, 2020

New effort with MyTrueAncestry.com

As a follow-up to the previous test I expanded tests including samples from a larger area.   These results look decent, although some individual results can be uncertain, which can be a result of bad aDna quality.   First my results and another Finnish result.  We both show similar Viking era history, but the difference exists in more ancient times.  The reason is probably that I am Western Finnish and she/he is Southwestern Finnish.   All samples are from academic sources (except mine) and randomly selected, so I can't assure that they represent any average for their ethnicity. 

The time of first results is limited to 1000 BC, but there is no particular reason to do it,  I just did it before I found out that at least in the south the decent history is much longer.  

My result:


Southwestern Finnish result:



Obviously she/he has more Eastern ancient origin than me, although the later history is very similar pointing to an Iron Age Scandinavian influence.  This conclusion is not far fetched, because my yDna reflects over 2000 years old Scandinavian origin and Southwest Finland got its Uralic population around 300-500 AD. 

For comparison a Swedish result.  It is obvious that her/his older history differs from what we see in Finnish results and is Central European (Saxon etc), until the Scandinavian ancestry stabilized in the Iron Age.






To see how the Fennoscandinavian ancestry differs from more eastern Uralic and Iron Age Steppe ancestry, I picked also a Mari and Tatar samples.


Mari:


Tatar:


And a Turkish sample, which shows decent ancient Anatolian ancestry:



Italian result from Tuscany:



I would say that there is quite a lot logic in these results, much more provable than not provable.  Even though there are ancestry points hard to explain, we can't say that even that ancestry is not always wrong.   Definitely there is a much potential to use aDna samples with reasonable algorithms to find out hidden ancestry and learn more about our history, plus dispel misconception. 

Thursday, July 9, 2020

Covid-19 risk alleles linked to the Neanderthal ancestry

According to Hugo Zeberg and Svante Pääbo there is a haplotype linked to Neanderthal ancestry and elevated risk of dying from Covid-19.  I tested 2504 samples from the 1000 genomes project and counted those minor alleles per population.  You can download results here.

edit. 09.07.20 20:30

Only one effective allele is included in commercial data and can be found in some personal results (like 23andme, FtDna etc):

rs10490770, allele C being a putative risk.  I also ran LD-analyses on the 1000genomes data, found several putative alleles, but didn't succeed to connect them to commercial data.  It is also good to notice that all listed risk alleles mostly form an unique haplotype. 

Abbreviation:


Sunday, June 21, 2020

My Mytrueancestry results

Before expanding my own analyses to this area I shortly publish something of my own results from MyTrueAncestry.  I know, here are many readers who don't give much credit to MTA and likely there are false and true results.  However, I have seen that those results are  usually in line with their customers' origins.

We see that my results focus on Estonian and Germanic aDna, but in my opinion the most interesting part of this is my older Trans-Volga Forest Steppe - Sintashta - Andronovo aDna.  Were  Trans-Volga and  Sintashta cultures the same and one culture speaking IE language, or parallel cultures,  the first speaking Uralic and the latter Indo-Iranian language?





Tuesday, May 5, 2020

Creating average genetic samples

Often especially ancient samples have poor quality, although sample number can be reasonable.  I made a simple Linux shell script making an average sample from a sample group, giving possibility to increase sample quality to a reasonable level for analyses based on allele frequencies.  The script reads EIGENSTRAT-format and the result is formed by picking alleles randomly from pooled samples.  Unfortunately Linux shell scripts don't support indexed files and I had to make some compromises to keep run time reasonable. The result of this script will not work with analyses based on IBD's or principal components, so it is not possible to use it f.ex. as an input file of the popular Eurogenes G25 test, but this should work with all analyses on Dodecad platform.  If someone is not familiar with these semantics it is very possible that the outcome is disappointing.  The script is freely available here.   I forgot, you need also a rs-id file and it is available here.  It should be unzipped to the same directory with the script.   Some results below using available academic samples and based on my own models:

Medieval Nomads

East_Asian 42.4
Uralic 24.5
Siberian 15.9
Northeast_European 6.1
Mediterranean 4.4
East_Scandinavian 3.3
Fennoscandinavian 2.2
AMBIG_European 1.2

Iberian Chalcolithic

Mediterranean 80.4
Northwest_European 15.6
Central_European 2.7
East_Scandinavian 1.4

Hungarian Bronze Age

Mediterranean 86.4
Central_European 11.3
Fennoscandinavian 1.1
East_Scandinavian 1.2

Polish Bronze Age

Northwest_European 39.0
Slavic 37.8
Fennoscandinavian 11.9
Central_European 6.5
East_Scandinavian 3.6
Baltic 1.2

Estonian Iron Age

Baltic 57.1
Slavic 26.1
Fennoscandinavian 11.5
Finnic 2.8
East_Scandinavian 2.2

edit 8.5. 15:30

The script edited so that it will accept also rs-id's in input, the original version accepted only  concatenation id's (chr:location).  Please notice that only hg19/GRCh37 mappings are possible.  New version is available here.