摘要 |
Systems, computer software and methods for calculating relatedness scores of domain names, which are indicative of relatedness of pairs of domain names requested by clients are described. The method includes receiving DNS traffic data, where the DNS traffic data includes at least domain names requested by the clients and identities of the clients requesting the domain names; generating, based on the identities of the clients, vectors including the requested domain names, where entries in the vectors correspond to client sessions in which the client has requested the domain names; reducing a dimensionality of the vectors by applying a dimensionality reduction method for generating reduced vectors; applying a similarity metric to the reduced vectors to calculate the relatedness scores; and storing the relatedness scores of the domain names.
|