Rapid Homoglyph Prediction and Detection

Rapid Homoglyph Prediction and Detection
复制标题

快速同形文字预测和检测

DOI:
--
复制
发表时间:
2018
期刊:
International Conference on Data Intelligence and Security
影响因子:
--
通讯作者:
Cui Yu
Cui Yu
中科院分区:
--
文献类型:
--
作者:
Avi Ginsberg;Cui Yu

文献摘要

被引文献

相似文献

随着技术渗透到地球仪,越来越需要支持越来越多的语言,因此也需要支持越来越多的字符集。许多字符集包含与其他字符集中的字符看起来相似或相同的字符(称为同型字符)。这带来了许多独特的问题,例如在网络钓鱼诈骗中经常使用的同源URL攻击。以前提出的解决这些问题的方案主要集中在限制用户可用的字符。本文提出了一种新的方法,可以预测和检测具有较高的准确率。这种方法模拟了人类的“视觉扫描”行为,并允许大量的视觉数据以一种易于快速搜索的格式表示。此外,它采用了预处理的时间记忆权衡,使检测和预测的homoglyphs在几分之一秒。
As technology permeates the globe, it is becoming necessary to support an increasing number of languages, and hence character sets. Many character sets contain characters that look similar or identical to characters from other character sets (known as homoglyph characters). This poses a number of unique problems, such as the homoglyph URL attacks frequently used in phishing scams. The previously proposed solutions to these problems have mostly focused on restricting the characters available to users. This paper proposes a new approach which can predict and detect homoglyphs with a high level of accuracy. This methodology simulates the human behavior of "visual scanning" and allows large amounts of visual data to be represented in a format that is easily and quickly searchable. Additionally, it employs a pre-processing time-memory trade-off to enable the detection and prediction of homoglyphs in fractions of a second.