Master'sOpen Access

Ottoman-Turkish optical character recognition and latin transcription

Is this your thesis?

This record came from a bulk archive import. If it’s yours, link it to your profile.

2016
0 views
0 downloads
Advisor: Doç. Dr. Fatih Koyuncu

Abstract (EN)

There are numerous documents in Ottoman-Turkish on the archives or online resources. Unfortunately these documents could not be understood by the people who cannot read Ottoman-Turkish alphabet. Ottoman-Turkish optical character recognition and Latin transcription could be the solution of this problem. In this thesis, Tesseract optical character recognition engine is used in order to recognize Ottoman-Turkish characters. Also, various methods are developed for the transcription from Ottoman Turkish to Latin. Characters on some Ottoman-Turkish images could not be recognized by optical character recognition methods. So, Ottoman-Turkish keyboard was developed for writing unrecognized characters with Ottoman-Turkish alphabet. Dictionary tables are used for transcription process. So enrichment data in the dictionary tables will increase of transcription success. Thus, an application was developed for enrichment data in the dictionary tables.

Author

Mustafa Doğru

How to Cite

Mustafa Doğru (Master Thesis). Ottoman-Turkish optical character recognition and latin transcription, 2016, Ankara Yıldırım Beyazıt University.

Keywords

License

Tüm Hakları Saklıdır

This work is shared under the specified license terms.

More theses from Ankara Yıldırım Beyazıt University