Automatic Processing of Historical Japanese Mathematics (Wasan) Documents

Yago Diez,Toya Suzuki,Marius Vila,Katsushi Waki

doi:10.3390/app11178050

Abstract

“Wasan” is the collective name given to a set of mathematical texts written in Japan in the Edo period (1603–1867). These documents represent a unique type of mathematics and amalgamate the mathematical knowledge of a time and place where major advances where reached. Due to these facts, Wasan documents are considered to be of great historical and cultural significance. This paper presents a fully automatic algorithmic process to first detect the kanji characters in Wasan documents and subsequently classify them using deep learning networks. We pay special attention to the results concerning one particular kanji character, the “ima” kanji, as it is of special importance for the interpretation of Wasan documents. As our database is made up of manual scans of real historical documents, it presents scanning artifacts in the form of image noise and page misalignment. First, we use two preprocessing steps to ameliorate these artifacts. Then we use three different blob detector algorithms to determine what parts of each image belong to kanji Characters. Finally, we use five deep learning networks to classify the detected kanji. All the steps of the pipeline are thoroughly evaluated, and several options are compared for the kanji detection and classification steps. As ancient kanji database are rare and often include relatively few images, we explore the possibility of using modern kanji databases for kanji classification. Experiments are run on a dataset containing 100 Wasan book pages. We compare the performance of three blob detector algorithms for kanji detection obtaining 79.60% success rate with 7.88% false positive detections. Furthermore, we study the performance of five well-known deep learning networks and obtain 99.75% classification accuracy for modern kanji and 90.4% for classical kanji. Finally, our full pipeline obtains 95% correct detection and classification of the “ima” kanji with 3% False positives.

Highlights

As it was impractical to run all parameter combinations with the full database, five Wasan images were used for the fine-tuning process, several parameter combinations where considered and the results for each of them was visually and automatically checked. minσ and maxσ are related to the size of the blobs detected, and we obtained the best results with values minσ ∈ (35, 45), maxσ ∈ (45, 50), the numσ parameter related to the step size when exploring the parameter space and was set to 5
4.3. “ima” Kanji Detection In Section 3.3, we presented a detailed study of how our algorithmic process can be used to locate an important structural element of Wasan Documents, the “ima” kanji that marks the location of the geometric problem diagram and is the first character in the textual description
The process described in this study is the first step in the construction of a database of Wasan documents

Summary

Introduction

The general public (samurai, merchants, farmers) learned a type of mathematics developed for people to enjoy, and Wasan documents have been used by Japanese citizens as a mathematics study tool or as a mental training hobby ever since [1,2]. These mathematics were studied from a completely different perspective from the mathematics learned in the West at the time and, in particular, did not have application goals [3]. These masters studied academic content at a level comparable to Western mathematics of the same period

Methods

Results

Discussion

Conclusion

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Automatic Processing of Historical Japanese Mathematics (Wasan) Documents

Abstract

Highlights

Summary

Talk to us

Similar Papers

More From: Applied sciences

Lead the way for us

Journal: Applied sciences	Publication Date: Aug 30, 2021
License type: CC BY 4.0

Similar Papers

Outcome Prediction of Postanoxic Coma: A Comparison of Automated Electroencephalography Analysis Methods
Stanley D T Pham ... Jeannette Hofmeijer
Neurocritical Care | VOL. 37
Stanley D T Pham, et. al.Stanley D T Pham ... Jeannette Hofmeijer
02 Mar 2022
Neurocritical Care | VOL. 37

Biological batch normalisation: How intrinsic plasticity improves learning in deep neural networks
Jeff Orchard ... Nolan Peter Shaw
-
Jeff Orchard, et. al.Jeff Orchard ... Nolan Peter Shaw
23 Sep 2020
23 Sep 2020

Biological batch normalisation: How intrinsic plasticity improves learning in deep neural networks.
Nolan Peter Shaw ... Tao Song
PloS one | VOL. 15
Nolan Peter Shaw, et. al.Nolan Peter Shaw ... Tao Song
23 Sep 2020
PloS one | VOL. 15

Identification framework for cracks on a steel structure surface by a restricted Boltzmann machines algorithm based on consumer-grade camera images
Yang Xu ... Shunlong Li
Structural Control and Health Monitoring | VOL. 25
Yang Xu, et. al.Yang Xu ... Shunlong Li
04 Aug 2017
Structural Control and Health Monitoring | VOL. 25

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Automatic Processing of Historical Japanese Mathematics (Wasan) Documents

Abstract

Highlights

Summary

Talk to us

Similar Papers

More From: Applied sciences