2016 COLING COLING 2016

Word Segmentation in Sanskrit Using Path Constrained Random Walks

Abstract

AbstractIn Sanskrit, the phonemes at the word boundaries undergo changes to form new phonemes through a process called as sandhi. A fused sentence can be segmented into multiple possible segmentations. We propose a word segmentation approach that predicts the most semantically valid segmentation for a given sentence. We treat the problem as a query expansion problem and use the path-constrained random walks framework to predict the correct segments.

🌉 Interdisciplinary Bridge — Interdisciplinary and Natural Language Processing
📈 Trend Setter — Morphology
🧭 Keyword Pioneer — semantic validity
🐣 Hot Topic Early Bird — word segmentation
🐝 Cross-Pollinator — Artificial Intelligence, Computer Science, Computer Vision, Data Science & Analytics, Deep Learning, Healthcare & Medicine, Interdisciplinary, Knowledge & Reasoning, Machine Learning, Mathematics & Optimization, Natural Language Processing, Reinforcement Learning, Security & Privacy, Speech & Audio