Augmented Data Selector to Initiate Text-Based CAPTCHA Attack

Ting Chen,Aolin Che,Hao Wang,Ke Zhang,Hong Xiao,Hong-Ning Dai,Yalin Liu

doi:10.1155/2021/9930608

Abstract

In the past decades, due to the low design cost and easy maintenance, text-based CAPTCHAs have been extensively used in constructing security mechanisms for user authentications. With the recent advances in machine/deep learning in recognizing CAPTCHA images, growing attack methods are presented to break text-based CAPTCHAs. These machine learning/deep learning-based attacks often rely on training models on massive volumes of training data. The poorly constructed CAPTCHA data also leads to low accuracy of attacks. To investigate this issue, we propose a simple, generic, and effective preprocessing approach to filter and enhance the original CAPTCHA data set so as to improve the accuracy of the previous attack methods. In particular, the proposed preprocessing approach consists of a data selector and a data augmentor. The data selector can automatically filter out a training data set with training significance. Meanwhile, the data augmentor uses four different image noises to generate different CAPTCHA images. The well-constructed CAPTCHA data set can better train deep learning models to further improve the accuracy rate. Extensive experiments demonstrate that the accuracy rates of five commonly used attack methods after combining our preprocessing approach are 2.62% to 8.31% higher than those without preprocessing approach. Moreover, we also discuss potential research directions for future work.

Highlights

Automated Public Turing tests to tell Computers and Humans Apart, abbreviated as CAPTCHA, is a kind of test that automatically distinguishes human and robot operations
We evaluate the proposed preprocessing approach by integrating it into five commonly used machine learning-based attack methods. ese commonly used attack methods include convolutional neural network (CNN), support vector machine (SVM), decision tree (DT), random forest (RF), and logistic regression (LR)
(2) We perform comprehensive ablation experiments by combining our approach in five commonly used machine/deep learning methods, including CNN, SVM, DT, RF, and LR. e experiment results confirm that our approach can significantly improve the high accuracy of the general attack methods, indicating the wide applicability of our approach

Summary

Introduction

Automated Public Turing tests to tell Computers and Humans Apart, abbreviated as CAPTCHA, is a kind of test that automatically distinguishes human and robot operations. Nearly 55% of websites use text-based CAPTCHAs as security and identification mechanisms, far exceeding other types, due to the low development and maintenance costs. For this reason, we mainly focus on the study of textbased CAPTCHAs in this paper. As an evolution of machine learning, deep learning is capable of making accurate decisions in the field of image recognition [12] For this reason, recent research efforts began to explore deep learning-based attack methods to solve text-based CAPTCHAs with a high-accurate accuracy rate [7, 8]. It is not worth wasting a lot of time for us to manually label these types of data

Objectives

Methods

Results

Discussion

Conclusion

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Security and Communication Networks	Publication Date: Jun 15, 2021
Citations: 3	License type: CC BY 4.0

R Discovery Prime

R Discovery Prime

Augmented Data Selector to Initiate Text-Based CAPTCHA Attack

Abstract

Highlights

Summary

Talk to us

Similar Papers

More From: Security and Communication Networks

Lead the way for us

Similar Papers

Counteracting Dark Web Text-Based CAPTCHA with Generative Adversarial Learning for Proactive Cyber Threat Intelligence
Ning Zhang ... Mohammadreza Ebrahimi
ACM Transactions on Management Information Systems | VOL. 13
Ning Zhang, et. al.Ning Zhang ... Mohammadreza Ebrahimi
10 Mar 2022
ACM Transactions on Management Information Systems | VOL. 13

Text-based CAPTCHA Vulnerability Assessment using a Deep Learning-based Solver
Daniel Aguilar ... Ricardo Flores Moyano
-
Daniel Aguilar, et. al.Daniel Aguilar ... Ricardo Flores Moyano
12 Oct 2021
12 Oct 2021

Make complex CAPTCHAs simple: A fast text captcha solver based on a small number of samples
Yao Wang ... Bailing Wang
Information Sciences | VOL. 578
Yao Wang, et. al.Yao Wang ... Bailing Wang
14 Jul 2021
Information Sciences | VOL. 578

Automating the Bypass of Image-based CAPTCHA and Assessing Security
Krish Sukhani ... Sarthak Maniar
-
Krish Sukhani, et. al.Krish Sukhani ... Sarthak Maniar
06 Jul 2021
06 Jul 2021

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Augmented Data Selector to Initiate Text-Based CAPTCHA Attack

Abstract

Highlights

Summary

Talk to us

Similar Papers

More From: Security and Communication Networks