A Deep Learning Architecture for Corpus Creation for Telugu Language

Dhana L. Rao, Venkatesh R.

doi:10.1007/978-981-15-4029-5_1

Citation Details

A Deep Learning Architecture for Corpus Creation for Telugu Language

Many natural languages are on the decline due to the dominance of English as the language of the World Wide Web (WWW), globalized economy, socioeconomic, and political factors. Computational Linguistics offers unprecedented opportunities for preserving and promoting natural languages. However, availability of corpora is essential for leveraging the Computational Linguistics techniques. Only a handful of languages have corpora of diverse genre while most languages are resource-poor from the perspective of the availability of machine-readable corpora. Telugu is one such language, which is the official language of two southern states in India. In this paper, we provide an overview of techniques for assessing language vitality/endangerment, describe existing resources for developing corpora for the Telugu language, discuss our approach to developing corpora, and present preliminary results. more »

Award ID(s):: 1730568

PAR ID:: 10174938

Author(s) / Creator(s):: Dhana L. Rao, Venkatesh R.

Date Published:: 2020-07-01

Journal Name:: Advances in intelligent systems and computing

Volume:: 1

ISSN:: 2194-5357

Page Range / eLocation ID:: 1-16

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Conference Paper:
https://doi.org/10.1007/978-981-15-4029-5_1

More Like this