科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Data in brief2026-08-01

BrajText-Saar: A structured cultural dataset for cultural text mining in indian heritage texts.

Mukta M Deshpande, Prafulla B Bafna

原始摘要(英文原文)· Original abstract
India is one of the countries where many regional languages are spoken in different parts of the country. Each part of the country has its unique cultural traditions; therefore, an abundance of cultural texts is available in Indian Regional languages. The cultural texts are mostly ancient scriptures that represent the community's heritage, values, and ancient knowledge. Braj is one of the Indian regional languages that represents the cultural heritage of the Braj region. The BrajText-Saar dataset contains texts from the Braj language, which features a variety of prehistoric scripts. The data was collected from the Maan Mandir trust's portal, which operates to maintain Braj culture and heritage. The Braj text primarily explains the divine actions and the life of Lord Krishna in the form of poetry and devotional songs. Offline Braj literature from renowned writers is available in manuscript form. This represents an opportunity for cultural text mining in the Braj Language. The developed dataset opens new paths in various research areas, including language studies, sentiment analysis, and emotion analysis, and is also suitable for digital humanities research.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

BrajText-Saar: A structured cultural dataset for cultural text mining in indian heritage texts. — 科研速览 Science Skim