paper-with-me

Papers

Abstractive Text Summarization for Contemporary Sanskrit Prose: Issues and Challenges

2025-01-03 · Shagun Sinha

This thesis presents Abstractive Text Summarization models for contemporary Sanskrit prose. The first chapter, titled Introduction, presents the motivation behind this work, the research questions, and the conceptual framework. Sanskrit is a low-resource inflectional language. The key research question that this thesis investigates is what the challenges in developing an abstractive TS for Sanskrit. To answer the key research questions, sub-questions based on four different themes have been posed in this work. The second chapter, Literature Review, surveys the previous works done. The third chapter, data preparation, answers the remaining three questions from the third theme. It reports the data collection and preprocessing challenges for both language model and summarization model trainings. The fourth chapter reports the training and inference of models and the results obtained therein. This research has initiated a pipeline for Sanskrit abstractive text summarization and has reported the challenges faced at every stage of the development. The research questions based on every theme have been answered to answer the key research question.

📄 PDF Abstract BibTeX arXiv:2501.01933

Code (0)

등록된 구현이 없습니다.

Tasks

Abstractive Text SummarizationLanguage ModelingLanguage ModellingText Summarization

Methods 이 논문이 사용한 방법론

TS Spatio-temporal features extraction that measure the stabilty. The proposed method is based on a compression algorithm named Run Length Encoding. The workflow of the method is…

Similar Papers 제목 키워드 기반

Abstractive Text Summarization for Sanskrit Prose: A Study of Methods and Approaches

2020-05-01 · LREC 2020 5 · Shagun Sinha, Girish Jha

The authors present a work-in-progress in the field of Abstractive Text Summarization (ATS) for Sanskrit Prose {--} a first attempt at ATS for Sanskrit (SATS). We will evaluate recent approaches and methods used for ATS …

Abstractive Text SummarizationExtractive Text SummarizationInformation RetrievalRetrieval+1

Sāmayik: A Benchmark and Dataset for English-Sanskrit Translation

2023-05-23 · Ayush Maheshwari, Ashim Gupta, Amrith Krishna, Atul Kumar Singh 외

We release S\={a}mayik, a dataset of around 53,000 parallel English-Sanskrit sentences, written in contemporary prose. Sanskrit is a classical language still in sustenance and has a rich documented heritage. However, due…

Machine TranslationTranslation

Still Not There: Can LLMs Outperform Smaller Task-Specific Seq2Seq Models on the Poetry-to-Prose Conversion Task?

2025-11-11 · Kunal Kingkar Das, Manoj Balaji Jagadeeshan, Nallani Chakravartula Sahith, Jivnesh Sandhan 외 arxiv

Large Language Models (LLMs) are increasingly treated as universal, general-purpose solutions across NLP tasks, particularly in English. But does this assumption hold for low-resource, morphologically rich languages such…

Survey on Abstractive Text Summarization: Dataset, Models, and Metrics

2024-12-22 · Gospel Ozioma Nnadi, Flavio Bertini

The advancements in deep learning, particularly the introduction of transformers, have been pivotal in enhancing various natural language processing (NLP) tasks. These include text-to-text applications such as machine tr…

Abstractive Text SummarizationGeneral KnowledgeImage to textMachine Translation+5

Poetry to Prose Conversion in Sanskrit as a Linearisation Task: A Case for Low-Resource Languages

2019-07-01 · ACL 2019 7 · Amrith Krishna, Vishnu Sharma, Bishal Santra, Aishik Chakraborty 외

The word ordering in a Sanskrit verse is often not aligned with its corresponding prose order. Conversion of the verse to its corresponding prose helps in better comprehension of the construction. Owing to the resource c…