2025
EMNLP
EMNLP 2025
PDFMathTranslate: Scientific Document Translation Preserving Layouts
Abstract
AbstractLanguage barriers in scientific documents hinder the diffusion and development of science and technologies. However, prior efforts in translating such documents largely overlooked the information in layouts. To bridge the gap, we introduce PDFMathTranslate, the world’s first open-source software for translating scientific documents while preserving layouts. Leveraging the most recent advances in large language models and precise layout detection, we contribute to the community with key improvements in precision, flexibility, and efficiency. The work is open-sourced at https://github.com/byaidu/pdfmathtranslate with more than 222k downloads.
🌉
Interdisciplinary Bridge
— Artificial Intelligence and Computer Vision and Natural Language Processing
🧭
Keyword Pioneer
— layout preservation
🐝
Cross-Pollinator
— Artificial Intelligence, Computer Science, Computer Vision, Data Science & Analytics, Deep Learning, Healthcare & Medicine, Interdisciplinary, Knowledge & Reasoning, Machine Learning, Mathematics & Optimization, Natural Language Processing, Reinforcement Learning, Robotics, Security & Privacy, Speech & Audio
Authors
Topics
Natural Language Processing > Generation > Text Generation
Natural Language Processing > Applications > Machine Translation
Natural Language Processing > Resources & Methods > Large Language Models
Natural Language Processing > Generation > Machine Translation
Computer Vision > Domain-Specific > Document Analysis
Computer Vision > Processing > Image Processing
Artificial Intelligence > Core AI > Language
Computer Vision > Applications > Document Analysis