PDFMathTranslate: Scientific Document Translation Preserving Layouts
Ouyang, Rongxin, Chu, Chang, Xin, Zhikuang, Ma, Xiangyao
–arXiv.org Artificial Intelligence
Language barriers in scientific documents hinder the diffusion and development of science and technologies. However, prior efforts in translating such documents largely overlooked the information in layouts. To bridge the gap, we introduce PDFMathTranslate, the world's first open-source software for translating scientific documents while preserving layouts. Leveraging the most recent advances in large language models and precise layout detection, we contribute to the community with key improvements in precision, flexibility, and efficiency. The work has been open-sourced at https://github.com/byaidu/pdfmathtranslate with more than 222k downloads.
arXiv.org Artificial Intelligence
Sep-23-2025
- Country:
- Asia
- China
- Myanmar > Chin State
- Hakha (0.04)
- Singapore > Central Region
- Singapore (0.04)
- Europe
- France (0.04)
- Germany (0.04)
- United Kingdom (0.04)
- North America > United States
- Illinois > Cook County > Chicago (0.04)
- Asia
- Genre:
- Research Report (0.64)
- Technology: