Phased haplotype features are used to infer an individual's ancestry. Reference genomic data is obtained for individuals of known ancestral origin. Haplotype features are identified based on consecutive SNPs from each individual. Sample genomic data is obtained for an individual of unknown ancestral origin. The data is phased and divided into features analogous to the features in the reference data. An admixture estimator then performs an admixture estimation based on the observed feature values in the sample data and the reference data. The estimation indicates a contribution of each of the known populations to the genome of the sample individual.
展开▼