MarSel : LD based tagSNP Selection System for Large-scale SNP Haplotype Dataset

MarSel : 대용량 SNP 일배체형 데이터에 대한 연관불균형기반의 tagSNP 선택 시스템

  • 김상준 (중앙대학교 컴퓨터공학부) ;
  • 여상수 (중앙대학교 컴퓨터공학부) ;
  • 김성권 (중앙대학교 컴퓨터공학부)
  • Published : 2006.02.01


Recently the tagSNP selection problem has been researched for reducing the cost of association studies between human's diversities and SNPs. General approach for this problem is that all of SNPs are separated into appropriate blocks and then tagSNPs are chosen in each block. Marsel in this paper is the system that involved the concept of linkage disequilibrium for overcoming the problem that the existing block partitioning approaches have short of biological meanings. In most approaches, the contiguous regions, which recombinations have LD coefficient |D'| and then tagSNP selection step is performed. And MarSel guarantees the minimum tagSNP selection using entropy-based optimal selection algorithm when tagSNPs are chosen in each block, and enables chromosome-level association studies using efficient memory management technique when input is very large-scale dataset that is impossible to be processed in the existing systems.


  1. J. I. Bell, 'Single Nucleotide Polymorphisms and Disease Gene Mapping,' Arthritis Research, Vol.4, pp.s273-s278, 2002
  2. R. C. Lewontin, 'The Interaction of Selection and Linkage. I. General Considerations; Heterotic Models,' Genetics, Vol.49, pp.49-67, 1964
  3. R. Mott, 'Marker Selection by Maximum Entropy,', Wellcome Trust Centre for Human Genetics, University of Oxford, 2003
  4. K. Zhang, Z. Qin, T. Chen, J. S.. Liu, M. S. Waterman and F. Sun, 'HapBlock: Haplotype Block Partitioning and Tag SNP Selection Software using a Set of Dynamic Programming Algorithms,' Bioinformatics, Vol.21(1), pp.131-134, 2003
  5. K. Zhang and L. Jin, 'HaploBlockFinder: Haplotype Block Analyses,' Bioinformatics, Vol.19, No.10, pp.1300-1301, 2003
  6. M. J. Daly, J. D. Rioux, S. F. Schaffner, T. J. Hudson and E. S. Lander, 'High-Resolution Haplotype Structure in the Human Genome,' Nature Genetics, Vol.29, No.2, pp.151-158, 2001
  7. N. Patil, A. J. Berno, D. A. Hinds, W. A. Barrett, J. M. Doshi, C. R. Hacker, C. R. Kautzer, D. H. Lee, C. Marjoribanks, D. P. McDonough, B. T. N. Nguyen, M. C. Norris, J. B. Sheehan, N. Shen, D. Stern, R. P. Stokowski, D. J Thomas, M. O. Trulson, K. R. Vyas, K. A. Frazer, S. P. A. Fodor, D. R. Cox, 'Blocks of Limited Haplotype Diversity Revealed by High-Resolution Scanning of Human Chromosome 21,' Science, Vol.294, pp.1719-1723, 2001
  8. K. Zhang, M. Deng, T. Chen, M. S. Waterman and F. Sun, 'A Dynamic Programming Algorithm For Haplotype Block Partitioning,' Proceedings of the National Academy of Sciences (PNAS), Vol.99, No.11, pp.7335-7339, 2002