Similarity analysis of DNA sequences through local distribution of nucleotides in strategic neighborhoods
This paper proposes a computationally efficient, alignment-free algorithm that represents DNA sequences as 24-dimensional vectors based on the local distribution of nucleotides in strategic neighborhoods, leveraging prime factorization uniqueness to achieve linear time complexity and low memory usage for effective phylogenetic analysis.